9router/open-sse/translator
whale9820 cfbdf06047 fix(volcengine-ark): clamp Kimi max_tokens to 32768 endpoint cap
VolcEngine Ark caps the Kimi family at max_tokens <= 32768, but the
model's advertised ceiling is far higher (Kimi-K2.7-Code resolves to
maxOutput 262144), so clampToModelMaxOutput alone leaves it uncapped and
the request 400s. Add a Kimi-scoped rule with an explicit maxOutputCap of
32768, combined with the model ceiling via min(). Covers max_tokens,
max_completion_tokens, max_output_tokens.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Co-authored-by: Cursor <cursoragent@cursor.com>
2026-07-09 15:10:05 +07:00
..
concerns fix(volcengine-ark): clamp Kimi max_tokens to 32768 endpoint cap 2026-07-09 15:10:05 +07:00
formats fix(claude): reconcile max_tokens vs thinking budget and lift per-model ceiling (#2381) 2026-07-05 17:38:16 +07:00
request fix(antigravity): align provider fingerprint with IDE Desktop 2.1.1 (#2389) 2026-07-08 10:17:47 +07:00
response feat(usage): track cached tokens + correct input/output/cache cost (#2209) 2026-07-03 15:18:27 +07:00
schema Refactor 2026-06-15 18:18:04 +07:00
formats.js feat: add CommandCode provider support 2026-05-07 23:01:33 +07:00
index.js fix(alicode): preserve cache_control for DashScope providers (#2069) 2026-06-29 15:29:49 +07:00