feat(opencode-go): expose Muse Spark reasoning effort aliases - #10883
diegosouzapw merged 2 commits into
Conversation
b580992
into
diegosouzapw:release/v3.8.50
|
Hi @excessivechaos @diegosouzapw, Testing with When a conversation history grows or has [400]: {"id":"chatcmpl_...","object":"chat.completion","model":"muse-spark-1.2-contributor-free","choices":[{"index":0,"message":{"role":"assistant"},"finish_reason":null}]}Observations:
|
|
Thanks for the detailed report @adevwithpurpose — we investigated this on a fresh deployment of the merged code and confirmed the upstream itself is rejecting the model, independent of this PR's changes. Direct upstream reproduction (no OmniRoute in the path)
So the empty-assistant 400 reproduces with the most minimal request possible — it is not gated on conversation history size, multi-turn, or effort configuration. This looks like an upstream serving defect (model listed in the catalog but the chat endpoint rejects it). This PR does not change the free-variant request path
Agreed follow-up (separate from the 400)Your point about context-window differentiation is valid and worth a separate fix: the free tier should advertise its ~128k ceiling for this model rather than inheriting the paid 1M context, and the empty- Happy to open a follow-up PR for the context-window/empty-response handling if that's wanted. |
…ouzapw#10883) Validado no worktree combinado do lote: typecheck:core, lint, gates de qualidade e testes focados (opencode-go-catalog-alignment + opencode-go-effort-aliases-8353, incluindo os novos casos muse-spark-1.2-contributor-*) todos verdes. CI vermelho neste PR é o base-red já rastreado em diegosouzapw#9985. Obrigado!
Summary
reasoning_effortat request timeVerification
The CommandCode effort work is intentionally separate.