VolcEngine Ark caps the Kimi family at max_tokens <= 32768, but the model's advertised ceiling is far higher (Kimi-K2.7-Code resolves to maxOutput 262144), so clampToModelMaxOutput alone leaves it uncapped and the request 400s. Add a Kimi-scoped rule with an explicit maxOutputCap of 32768, combined with the model ceiling via min(). Covers max_tokens, max_completion_tokens, max_output_tokens. Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com> Co-authored-by: Cursor <cursoragent@cursor.com> |
||
|---|---|---|
| .. | ||
| config | ||
| executors | ||
| handlers | ||
| providers | ||
| rtk | ||
| services | ||
| shared | ||
| transformer | ||
| translator | ||
| utils | ||
| .npmignore | ||
| AGENTS.md | ||
| index.js | ||