mirror of
https://github.com/Nezumi-2711/9router.git
synced 2026-09-22 13:38:31 +00:00
VolcEngine Ark caps the Kimi family at max_tokens <= 32768, but the model's advertised ceiling is far higher (Kimi-K2.7-Code resolves to maxOutput 262144), so clampToModelMaxOutput alone leaves it uncapped and the request 400s. Add a Kimi-scoped rule with an explicit maxOutputCap of 32768, combined with the model ceiling via min(). Covers max_tokens, max_completion_tokens, max_output_tokens. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com> Co-authored-by: Cursor <cursoragent@cursor.com>