Gemini CLI requests with small max_tokens spend the whole output budget on thoughts after reasoning_effort maps to thinkingConfig, returning blank content or finish=length. Raise maxOutputTokens floors per thinking level/ budget (clamped to caps.maxOutput). Also emit toolConfig functionCallingConfig.mode=VALIDATED for Gemini CLI tool requests to avoid MALFORMED_FUNCTION_CALL. Co-authored-by: Cursor <cursoragent@cursor.com> |
||
|---|---|---|
| .. | ||
| config | ||
| executors | ||
| handlers | ||
| providers | ||
| rtk | ||
| services | ||
| shared | ||
| transformer | ||
| translator | ||
| utils | ||
| .npmignore | ||
| AGENTS.md | ||
| index.js | ||