thienpv and Cursor
288940960a
fix(translator): clamp thinking effort max->xhigh for OpenAI format ( #2466 )
...
Claude Code sends reasoning_effort "max" (its top level); OpenAI enum caps
at "xhigh" and rejects "max" with HTTP 400 "max effort not support".
applyFormat case "openai" now clamps "max"->"xhigh" before assigning
body.reasoning_effort; other levels pass through unchanged.
Add regression test covering client output_config.effort, direct
reasoning_effort, passthrough of xhigh/high, and budget_tokens capping.
Co-authored-by: Cursor <cursoragent@cursor.com >
2026-07-09 15:13:24 +07:00
decolua
b10b807063
# v0.5.20 (2026-07-07)
...
## Features
- **Thinking**: per-model thinking level picker on provider page — appends `(level)` suffix to copied model names for forced reasoning effort across all formats (openai, claude, gemini, deepseek, kimi, qwen, zai, minimax, hunyuan, step)
- **RTK**: add JS-native git-log filter (#2423 )
- **Caveman**: add targeted upstream-aligned style rules (#2424 )
- **i18n**: add Farsi (fa) language support (#2385 )
## Fixes
- **Thinking**: strip `(level)` suffix from upstream `body.model` so providers no longer reject requests
- **Translator**: preserve developer instructions in openai-responses conversion (#2434 )
- **count_tokens**: count structured Anthropic blocks (#2419 )
- **Volcengine-ark**: clamp GLM-5 max_tokens to model output ceiling (#2428 )
- **Kimi**: normalize reasoning_effort to backend enum (#2427 )
- **Claude**: reconcile max_tokens vs thinking budget and lift per-model ceiling (#2381 )
- **Kiro**: deliver system prompt natively, add Opus 4.5/4.7/4.8, tolerate dash version ids (#2366 )
- **Headroom**: proxy dashboard through app (#2372 )
- **MITM**: recover from stale lock file on server start
2026-07-07 16:29:11 +07:00
whale and Cursor
8c068a1f5c
fix(kimi): normalize reasoning_effort to backend enum ( #2427 )
...
Map auto→high, minimal→low, xhigh→max and whitelist low/medium/high/max
so Kimi/kimchi SGLang backends no longer receive invalid effort values.
Co-authored-by: Cursor <cursoragent@cursor.com >
2026-07-07 11:56:39 +07:00
ntdung6868 and Cursor
3a866fe18d
fix(reasoning): preserve effort through Codex translations
...
Carry Claude reasoning_effort/reasoning into OpenAI Chat, map into
OpenAI Responses reasoning.effort, and keep request-level effort
(incl. xhigh) across tool-result turns instead of collapsing to high.
Co-authored-by: Cursor <cursoragent@cursor.com >
2026-06-26 12:06:35 +07:00
Mink Nguyen and Cursor
c4f80d30d8
fix provider thinking compatibility
...
- claude: handle DeepSeek thinking blocks defensively, unsigned placeholder; fix kept-vs-seen thinking detection
- gemini: clamp unsupported max/xhigh thinking levels to high
- testUtils: probe Cloud Code Assist for gemini-cli/antigravity with 401 refresh retry
- tests: add translator regression coverage
Co-authored-by: Cursor <cursoragent@cursor.com >
2026-06-26 10:12:48 +07:00
decolua
d03f9fb823
Enhance configuration and model capabilities
2026-06-16 23:32:28 +07:00
decolua
b282f05549
Refactor
2026-06-15 18:18:04 +07:00