Repository navigation
feat(models): declare reasoning efforts for gemini - #3349
Merged
Merged
Conversation
Contributor
WalkthroughFive Google AI Studio and Google Vertex Gemini configurations now declare support for minimal, low, medium, and high reasoning effort levels. ChangesGoogle Gemini reasoning effort support
Estimated code review effort: 1 (Trivial) | ~2 minutes Possibly related PRs
🚥 Pre-merge checks | ✅ 5✅ Passed checks (5 passed)
✨ Finishing Touches🧪 Generate unit tests (beta)
Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out. Comment |
gemini-3.1-flash-lite declared reasoningEfforts on its google-vertex mapping but not on google-ai-studio, unlike the other three Gemini models in this change which declare the same tiers on both providers. Verified live against AI Studio: minimal/low/medium/high all return 200 with reasoning output. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
steebchen
enabled auto-merge
August 3, 2026 16:55
steebchen
added this pull request to the merge queue
Aug 3, 2026
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Declares
reasoningEffortsfor the Gemini 3.x mappings that were missing them, per #3028.Changes
google-vertex:gemini-3.1-flash-lite,gemini-3.5-flash,gemini-3.5-flash-lite,gemini-3.6-flash→["minimal", "low", "medium", "high"]google-ai-studio:gemini-3.1-flash-lite→["minimal", "low", "medium", "high"](its Vertex sibling declared the tiers but the AI Studio mapping did not, unlike the other three models which declare them on both providers)Source: Google Cloud's Vertex Thinking docs — summary table lists exact supported
thinking_levelvalues per model, cross-checked against each model's individual Vertex docs page.Note that the gateway does not forward
thinking_levelitself; for Google providers it translates each effort tier into athinkingConfig.thinkingBudgetvalue (prepare-request-body.ts). The declared tiers are what the catalogue publishes and what the e2e suite exercises.Not changed — with reasons
thinking_levelmechanism exists for these generations (thinking_budget, a raw token count, is used instead — already declared viareasoningMaxTokens).gemini-3-pro-preview,gemini-3.1-flash-lite-preview— retired/deactivated on Vertex.gemini-3-flash-preview,gemini-3.1-pro-preview— the Vertex thinking table has rows for "Gemini 3 Flash"/"Gemini 3.1 Pro" without "preview" in the name; couldn't confirm these are the same model as our catalog's preview-suffixed entries (dedicated per-model doc pages both 404). Left unset rather than assume.none— not declared anywhere here, matching the sibling mappings. It does return 200 on AI Studio with thinking disabled, but declaring it is out of scope for this change.Testing
pnpm format,pnpm build— clean.FULL_MODE=true, which expands one case per declared tier):google-vertex× 4 models: 16/16 effort cases pass (minimal/low/medium/high).google-ai-studio/gemini-3.1-flash-lite: 4/4 effort cases pass; full scoped run 113 passed / 0 failed.gemini-3.1-flash-liteon Vertex: 140 → 152 → 243 → 271 for minimal → low → medium → high), matching Vertex's ownthoughtsTokenCountinusageMetadata.Summary by CodeRabbit