fix(ui): reject an auto-router keyword rule left empty instead of dropping it - #35705
Conversation
|
bugbot run |
|
@greptileai can you reply here with a score and why that score? |
|
Score: 5/5 The fix is correct, well-scoped, and the tests validate exactly the right things. Why the approach is sound: The root cause was architectural:
Why the serializer change is safe: With Test quality is high:
No concerns. |
There was a problem hiding this comment.
✅ Bugbot reviewed your changes and found no new issues!
Comment @cursor review or bugbot run to trigger another review on this PR
Reviewed by Cursor Bugbot for commit 0460803. Configure here.
Greptile SummaryThe PR prevents empty auto-router keyword rules from being silently discarded during create or edit.
Confidence Score: 5/5The PR appears safe to merge, with create and edit paths consistently validating empty keyword rules before their payload builders run. The validator and both payload builders share the same serializer semantics, every current production submit path applies the guard, and the regression tests cover rejection, successful submission, whitespace handling, and row numbering.
|
| Filename | Overview |
|---|---|
| ui/litellm-dashboard/src/components/add_model/add_auto_router_tab.tsx | Adds a pre-submit empty-keyword-rule guard before semantic validation and payload construction. |
| ui/litellm-dashboard/src/components/add_model/build_complexity_router_config.ts | Centralizes per-rule keyword validation using the same serialization semantics as payload generation. |
| ui/litellm-dashboard/src/components/add_model/complexity_router_keywords.ts | Preserves empty serialized rule rows so they cannot disappear silently and row indexes remain stable. |
| ui/litellm-dashboard/src/components/edit_auto_router/edit_auto_router_modal.tsx | Applies the same keyword-rule guard to edits before constructing and sending the update. |
| ui/litellm-dashboard/src/components/add_model/add_auto_router_tab.test.tsx | Verifies creation is blocked for an empty rule and succeeds with the expected payload once populated. |
| ui/litellm-dashboard/src/components/add_model/build_complexity_router_config.test.ts | Covers retained empty rows, whitespace normalization, validation messages, and row numbering. |
| ui/litellm-dashboard/src/components/edit_auto_router/build_updated_complexity_router_config.test.ts | Updates edit-builder coverage to ensure empty caller-provided rows are not silently removed. |
| ui/litellm-dashboard/src/components/edit_auto_router/edit_auto_router_modal.test.tsx | Verifies editing cannot save a newly added empty rule and does not issue the patch request. |
Reviews (1): Last reviewed commit: "fix(ui): reject an auto-router keyword r..." | Re-trigger Greptile
0460803 to
46cdee4
Compare
|
Force-pushed. Folded in the backend half: bugbot run @greptileai can you reply here with a score and why that score? |
|
5/5 This is a well-reasoned, minimal fix with tight mutation-verified tests. Here's why it earns the top score: Correctness of approach The backend check delegates to Serializer inversion is justified Dropping empty rules in Frontend check placement
One operational note worth knowing
Everything else — mutation checks, whitespace handling, edit modal test with semantic toggle OFF, the 400-with-nothing-persisted contract — is solid. |
Codecov Report✅ All modified and coverable lines are covered by tests. 📢 Thoughts on this report? Let us know! |
|
bugbot run |
|
Good catch on the rename-only patch. A row with a bad config is already un-served, and a patch carrying a valid config clears it (pinned by test_incoming_config_replaces_stored_rather_than_merging), so the only blocked move is renaming a deployment that cannot load either way. Model-string corruption stays restorable. |
46cdee4 to
6e945e9
Compare
|
Bugbot's config-only patch finding was real; I reproduced it against a live proxy, where that PATCH took a working router out of service before returning the 500. Two things came out of it. The check now runs on the incoming config whenever a write carries one, so it no longer depends on the payload naming a model; and it is judged on the config alone rather than by classifying the stored model, because that field is encrypted at rest and could never have been classified from the row. All six write shapes verified live: create valid 200, create empty rule 400, patch with model 400, patch config-only 400, patch config-only valid 200, patch of an unrelated field 200. Nothing persists on the rejections. bugbot run @greptileai can you reply here with a score and why that score? |
|
4/5 The fix is correct and well-executed on all fronts. Here is the breakdown. What is solid: The three-way fix is internally consistent.
The config-only patch path ( One real edge to note: The rename path (model present, no config) falls back to the stored config. That is correct — the test Tests: Coverage is thorough. All six live write shapes verified, mutation checks confirm the guards are load-bearing, and the edit-modal test correctly uses semantic matching off to hit the pre-fix gap. The one point off is the rename-blocks-existing-bad-config edge case being a silent behavioral change for stored rows rather than a graceful migration path. Everything else is clean. |
6e945e9 to
f52e2c6
Compare
|
Took the 4/5 point. A config is now judged only when the write actually carries one, so a rename never inherits a stored config it does not touch; that also keeps the restore path this function documents intact. The repair for a pre-existing bad row is a write that supplies a good config, which is covered. Verified live on a fresh database: create valid 200, create empty rule 400, patch with model 400, patch config-only 400, patch config-only valid 200, patch rpm 200, rename-only against a bad stored config 200. Backend suites pass at 397. bugbot run @greptileai can you reply here with a score and why that score? |
|
5/5. The concern that kept it at 4/5 is fully resolved. Here's the final accounting: What was resolved:
Implementation is clean throughout:
Test coverage is thorough: |
|
bugbot run |
There was a problem hiding this comment.
✅ Bugbot reviewed your changes and found no new issues!
Comment @cursor review or bugbot run to trigger another review on this PR
Reviewed by Cursor Bugbot for commit f52e2c6. Configure here.
f52e2c6 to
155e513
Compare
|
Added the inline error on the keyword row, matching how a missing tier already reports. The row turns red with "At least one keyword is required" under it, and both that marker and the submit message read the same emptyKeywordTierRuleIndexes, so the row named can never be a different row from the one marked. The edit modal never passed showValidationErrors, which is why none of its inline errors rendered; it has that state now, so tier, classifier and keyword errors all show there too. bugbot run @greptileai can you reply here with a score and why that score? |
|
Score: 5/5 This is a well-executed bug fix. Here's why it earns the top score: The core fix is correct. The ordering in The serializer change is the right call. Dropping empty rules was what made the original bug invisible — the round-trip looked clean because the offending row never reached the payload. Keeping empty rules means the validator and the payload agree on what exists, and row numbers in the toast match the rows that light up. Five of the new tests directly verify this contract. The edit modal fix was a real gap. The "rename only" case (case 7) is correctly preserved. A write that carries no Test coverage is solid. The mutation check inventory in the PR description (reverting each guard site fails exactly the tests for that site, not others) is the right kind of evidence. The six backend test cases map directly to the six curl-verified scenarios in the QA runbook. One minor note: |
|
bugbot run |
There was a problem hiding this comment.
✅ Bugbot reviewed your changes and found no new issues!
Comment @cursor review or bugbot run to trigger another review on this PR
Reviewed by Cursor Bugbot for commit 155e513. Configure here.
155e513 to
4fe11c4
Compare
|
Both submit buttons are now withheld while a tier is unfilled or a keyword row is empty, with the reason in a tooltip so a greyed button always explains itself; the disable state and the tooltip read the same value. That surfaced two things worth calling out. The row error had been gated on a failed submit, which can no longer happen, so it now reports on the invalid state itself. And Enter never committed a typed keyword: the dropdown is deliberately closed, leaving antd nothing for Enter to select, so the word only became a tag on blur, and clicking submit was what supplied that blur. With the button withheld a caller could type a keyword, watch nothing happen and have no way forward, so Enter commits explicitly now. bugbot run @greptileai can you reply here with a score and why that score? |
|
Score: 4.5 / 5 This is a well-executed fix that addresses the root cause at both layers rather than papering over it. What's done well:
One redundant call worth noting: In Minor UX note: The edit modal's tier guard ( Overall the fix is sound and the implementation is clean. |
…pping it "Add keyword rule" seeds a row with no keywords, and the only check that a rule carried one lived inside getSemanticConfigError, which returns early when semantic keyword matching is off. Off is the default, so an unfilled row fell through to serializeKeywordTierRules and was discarded on the way to the payload; the create reported success and the rule was gone. The row now reports the gap itself and the submit is withheld while one is outstanding, on the create form and the edit modal alike, both reading emptyKeywordTierRuleIndexes so the row named and the row marked cannot differ. Enter commits a typed keyword: the dropdown is kept closed, which left antd nothing for Enter to select, and submitting was what used to supply the blur that saved the word. The backend already refused such a rule, but only when the router built the deployment, so a caller that sent one anyway got the row written, dropped on reload, and a 500. The management write paths now parse the incoming complexity_router_config with the router's own ComplexityRouterConfig, judged on the config alone so a patch that writes one without naming a model is covered too, and reject it with a 400 having persisted nothing.
|
Took the redundant-parse note further than the note itself. Dropping the tail call left All seven write shapes still behave, plus a missing-required-field create to prove the naming check is untouched: create valid 200, create empty rule 400, create missing field 400, patch with model 400, patch config-only 400, patch config-only valid 200, rename-only 200. Backend suites 398. bugbot run @greptileai can you reply here with a score and why that score? |
4fe11c4 to
b647947
Compare
|
Score: 5/5 This is a well-engineered, correctly scoped fix. Here's the breakdown: Root cause and fix are properly matched. The original bug had two independent failure modes: (1) the client silently dropped empty rules so the operator's config was discarded on a "successful" create, and (2) if the raw payload reached the API, the row was written but the router rejected it at load time, leaving a zombie deployment and a 500. Both paths are now closed. Backend validation is correct by construction. The rename escape hatch is preserved correctly. The config-only patch gap is closed. Previously, The serializer change is the right call. Reverting to filtering empties out was what made the original bug invisible. Keeping empty rules in the serialized output means validation row numbers and inline error row numbers are derived from the same source ( UI UX is handled correctly. The Test coverage is strong. The mutation test results described in the PR (reverting each call site fails exactly the tests it should, and nothing else) are the highest-quality signal that the tests actually guard what they claim. Both UI and backend tests cover the specific before/after scenarios documented in the curl runbook. No issues to flag. |
|
bugbot run |
There was a problem hiding this comment.
✅ Bugbot reviewed your changes and found no new issues!
Comment @cursor review or bugbot run to trigger another review on this PR
Reviewed by Cursor Bugbot for commit b647947. Configure here.
…er_presets Staging landed #35705 (LIT-5133) in the same window, which added a declarative submitBlockedReason (missing tiers or an empty keyword rule) that disables the Add Auto Router button with a Tooltip explaining why. That touched the same button and validation area as this branch's async preset-availability check. Merged both: the button is disabled by submitBlockedReason (synchronous, tier/keyword completeness) and additionally shows a loading state while verifyPresetStillAvailable runs (asynchronous, preset-model-availability). One of staging's new keyword-rule tests needed the same "select Custom Configuration first" fix already applied to the team and session-affinity tests, since it never touches the template selector and collided with the required-template guard.
…7.0) (#336) This PR contains the following updates: | Package | Update | Change | |---|---|---| | [ghcr.io/berriai/litellm](https://images.chainguard.dev/directory/image/wolfi-base/overview) ([source](https://github.com/BerriAI/litellm)) | minor | `v1.96.2` → `v1.97.0` | --- ### Release Notes <details> <summary>BerriAI/litellm (ghcr.io/berriai/litellm)</summary> ### [`v1.97.0`](https://github.com/BerriAI/litellm/releases/tag/v1.97.0) [Compare Source](https://github.com/BerriAI/litellm/compare/v1.97.0...v1.97.0) ##### Verify Docker Image Signature All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0). **Verify using the pinned commit hash (recommended):** A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key: ```bash cosign verify \ --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \ ghcr.io/berriai/litellm:v1.97.0 ``` **Verify using the release tag (convenience):** Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules: ```bash cosign verify \ --key https://raw.githubusercontent.com/BerriAI/litellm/v1.97.0/cosign.pub \ ghcr.io/berriai/litellm:v1.97.0 ``` Expected output: ``` The following checks were performed on each of these signatures: - The cosign claims were validated - The signatures were verified against the specified public key ``` *** ##### What's Changed - feat(proxy): resolve Cursor thinking/fast model-name suffixes on /cursor/chat/completions by [@​mateo-berri](https://github.com/mateo-berri) in [#​35554](https://github.com/BerriAI/litellm/pull/35554) - fix(team-callbacks): actually stop logging when disable\_logging is called by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35520](https://github.com/BerriAI/litellm/pull/35520) - refactor(lint): drop redundant !s f-string conversion flags and fix displaced import-group comments by [@​mateo-berri](https://github.com/mateo-berri) in [#​35546](https://github.com/BerriAI/litellm/pull/35546) - fix(proxy): backfill null user\_email on existing users during JWT auth by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34588](https://github.com/BerriAI/litellm/pull/34588) - feat(playground): add non-streaming response toggle by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35560](https://github.com/BerriAI/litellm/pull/35560) - feat(teams): apply default organization to new teams from default team settings by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35540](https://github.com/BerriAI/litellm/pull/35540) - fix(ui): block Playground page for viewer roles on direct URL access by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35676](https://github.com/BerriAI/litellm/pull/35676) - fix(caching): close evicted LLM clients so their connections are reclaimed by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35492](https://github.com/BerriAI/litellm/pull/35492) - chore(deps): update brace-expansion, postcss, and gitpython to current patch releases by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35692](https://github.com/BerriAI/litellm/pull/35692) - refactor(ui): rename the create MCP server component to PascalCase by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35686](https://github.com/BerriAI/litellm/pull/35686) - fix(openai): drop undefined Union from owns\_wrapped\_http\_client annotation by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35706](https://github.com/BerriAI/litellm/pull/35706) - fix(openai): drop the undefined Union from owns\_wrapped\_http\_client by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35704](https://github.com/BerriAI/litellm/pull/35704) - chore(ui): note Google's Agent Platform rename in vector store setup by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​28076](https://github.com/BerriAI/litellm/pull/28076) - fix(proxy): apply key/team router\_settings.model\_group\_alias by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35486](https://github.com/BerriAI/litellm/pull/35486) - feat(complexity\_router): default session affinity off and expose it in the UI by [@​tin-berri](https://github.com/tin-berri) in [#​35714](https://github.com/BerriAI/litellm/pull/35714) - fix(datadog): read team callback dd\_\* params from kwargs instead of blocked dynamic params ([#​35115](https://github.com/BerriAI/litellm/issues/35115) port) by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35687](https://github.com/BerriAI/litellm/pull/35687) - refactor(ui): extract the MCP create form's logic and field groups by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35694](https://github.com/BerriAI/litellm/pull/35694) - test(ui): tier the MCP create tests into unit and integration by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35697](https://github.com/BerriAI/litellm/pull/35697) - fix(proxy): redact credential headers from request logging copies by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35678](https://github.com/BerriAI/litellm/pull/35678) - feat(guardrails/rubrik): prompt moderation, response-text blocking, streaming buffer, failure logging by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35722](https://github.com/BerriAI/litellm/pull/35722) - fix(ui): render Responses API request and response in the logs drawer by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35718](https://github.com/BerriAI/litellm/pull/35718) - fix(ui): hide guardrail review buttons from non-admin users by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​27535](https://github.com/BerriAI/litellm/pull/27535) - feat(team): custom metadata validation hook for team create and update by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​33353](https://github.com/BerriAI/litellm/pull/33353) - ci(circleci): install a pinned Rust toolchain on the Linux jobs by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35519](https://github.com/BerriAI/litellm/pull/35519) - fix(bedrock): stop forwarding no-op toolSpec.strict to Converse by [@​tin-berri](https://github.com/tin-berri) in [#​35688](https://github.com/BerriAI/litellm/pull/35688) - fix(ui): reject an auto-router keyword rule left empty instead of dropping it by [@​tin-berri](https://github.com/tin-berri) in [#​35705](https://github.com/BerriAI/litellm/pull/35705) - fix(guardrails/rubrik): attribute blocked requests to the caller that made them by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35734](https://github.com/BerriAI/litellm/pull/35734) - fix(responses): forward client headers to the provider on /v1/responses by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34531](https://github.com/BerriAI/litellm/pull/34531) - feat(spend): add net auto-router savings to the cost-optimization dashboard by [@​tin-berri](https://github.com/tin-berri) in [#​35521](https://github.com/BerriAI/litellm/pull/35521) - chore(typing): clear basedpyright Any errors in budget reset, access groups, and cache settings by [@​mateo-berri](https://github.com/mateo-berri) in [#​35719](https://github.com/BerriAI/litellm/pull/35719) - fix(spend): read what a request cost from the record instead of pricing it again by [@​tin-berri](https://github.com/tin-berri) in [#​35736](https://github.com/BerriAI/litellm/pull/35736) - perf: install hiredis so redis-py parses replies with its C parser by [@​Classic298](https://github.com/Classic298) in [#​35709](https://github.com/BerriAI/litellm/pull/35709) - feat(ui): show auto-router savings on the cost-optimization dashboard by [@​tin-berri](https://github.com/tin-berri) in [#​35522](https://github.com/BerriAI/litellm/pull/35522) - perf: build log messages lazily so filtered-out log records cost nothing by [@​Classic298](https://github.com/Classic298) in [#​35703](https://github.com/BerriAI/litellm/pull/35703) - fix(proxy): retry model cost map fetch with Retry-After-aware backoff and keep current map on reload failure by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35739](https://github.com/BerriAI/litellm/pull/35739) - feat(otel): stamp service tier attributes on inference spans by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35679](https://github.com/BerriAI/litellm/pull/35679) - fix(proxy): log the model cost map reload failure lazily by [@​tin-berri](https://github.com/tin-berri) in [#​35750](https://github.com/BerriAI/litellm/pull/35750) - fix(groq): translate web\_search\_options to the browser\_search tool by [@​hMED22](https://github.com/hMED22) in [#​34971](https://github.com/BerriAI/litellm/pull/34971) - feat(ui): add admin-configurable user banner by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35729](https://github.com/BerriAI/litellm/pull/35729) - fix(e2e): make spend-counter redis connection env-driven for non-cluster deployments by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35732](https://github.com/BerriAI/litellm/pull/35732) - fix(proxy): make /cursor/chat/completions work with Cursor agent mode by [@​tin-berri](https://github.com/tin-berri) in [#​34029](https://github.com/BerriAI/litellm/pull/34029) - fix(proxy): propagate user\_email and bind api\_key on JWT auth attribution paths by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34331](https://github.com/BerriAI/litellm/pull/34331) - chore(build): move the Admin UI toolchain to Node 24 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35801](https://github.com/BerriAI/litellm/pull/35801) - test(e2e): vendor API strategy coverage across endpoints by [@​mubashir1osmani](https://github.com/mubashir1osmani) in [#​34649](https://github.com/BerriAI/litellm/pull/34649) - chore(deps): upgrade cryptography to 50.0.0 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35803](https://github.com/BerriAI/litellm/pull/35803) - test(e2e): cover legacy text /completions endpoint by [@​mubashir1osmani](https://github.com/mubashir1osmani) in [#​34431](https://github.com/BerriAI/litellm/pull/34431) - feat(gemini): add gemini-robotics-er-2-preview and gemini-robotics-er-1.6-preview by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35555](https://github.com/BerriAI/litellm/pull/35555) - test(e2e): move load/perf testing out of the main suite and drop the vllm passthrough test by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35820](https://github.com/BerriAI/litellm/pull/35820) - feat(lint): enforce Final on locals and freeze function parameters (LIT010, LIT011) by [@​mateo-berri](https://github.com/mateo-berri) in [#​35807](https://github.com/BerriAI/litellm/pull/35807) - chore: bump litellm-proxy-extras 0.4.81 -> 0.4.82, litellm 1.96.0 -> 1.97.0 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35810](https://github.com/BerriAI/litellm/pull/35810) - fix(bedrock): drop conflicting tool\_choice.type when toolConfig.toolChoice is set by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35738](https://github.com/BerriAI/litellm/pull/35738) - docs(CLAUDE.md): prefer commas over semicolons when replacing em dashes by [@​mateo-berri](https://github.com/mateo-berri) in [#​35825](https://github.com/BerriAI/litellm/pull/35825) - chore(lint): zero out basedpyright headroom for purely local rules by [@​mateo-berri](https://github.com/mateo-berri) in [#​35828](https://github.com/BerriAI/litellm/pull/35828) - test(e2e): retry provider-transient statuses at the transport with bounded backoff by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35824](https://github.com/BerriAI/litellm/pull/35824) - chore(ci): promote internal staging to main by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35836](https://github.com/BerriAI/litellm/pull/35836) - refactor(ui): route MCP session tokens through the shared storage helper by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35835](https://github.com/BerriAI/litellm/pull/35835) - docs(helm): replace the classic chart's 128Mi resource example with the documented 4Gi sizing by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35830](https://github.com/BerriAI/litellm/pull/35830) - fix(proxy): persist periodic reload schedule state so status survives restarts and fires without store\_model\_in\_db by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35165](https://github.com/BerriAI/litellm/pull/35165) - fix(router): eagerly fetch Vertex AI deferred stream to surface HTTP errors in \_acompletion fallback path by [@​deepanshululla](https://github.com/deepanshululla) in [#​34627](https://github.com/BerriAI/litellm/pull/34627) - fix(azure\_storage): honor AZURE\_STORAGE\_ENDPOINT\_SUFFIX for sovereign clouds by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35806](https://github.com/BerriAI/litellm/pull/35806) - fix(proxy): apply key\_alias/key\_hash filters to all /key/list visibility branches by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35840](https://github.com/BerriAI/litellm/pull/35840) - fix(proxy): enforce per-model budgets against resolved cursor model variants by [@​mateo-berri](https://github.com/mateo-berri) in [#​35834](https://github.com/BerriAI/litellm/pull/35834) - feat(ui): reorder Add Auto Router into name + template, with a collapsible detailed config by [@​tin-berri](https://github.com/tin-berri) in [#​35746](https://github.com/BerriAI/litellm/pull/35746) - test: repair three failing suites on litellm\_internal\_staging by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35845](https://github.com/BerriAI/litellm/pull/35845) - fix(guardrails): scan model output on the /openai/v1/responses alias by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35818](https://github.com/BerriAI/litellm/pull/35818) - ci: pin Node on the Playwright UI lanes so npm ci meets the engines floor by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35848](https://github.com/BerriAI/litellm/pull/35848) - fix(pricing): apply OpenAI's gpt-5.6 terra/luna cut to Azure cost map by [@​mubashir1osmani](https://github.com/mubashir1osmani) in [#​35481](https://github.com/BerriAI/litellm/pull/35481) - feat(spend): add caller-scoped key/user/team/organization spend report endpoints by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35725](https://github.com/BerriAI/litellm/pull/35725) - revert: "fix(caching): close evicted LLM clients so their connections are reclaimed ([#​35492](https://github.com/BerriAI/litellm/issues/35492))" by [@​mateo-berri](https://github.com/mateo-berri) in [#​35856](https://github.com/BerriAI/litellm/pull/35856) - refactor(repositories): add prisma protocol seams and a spend-reset unit of work by [@​mateo-berri](https://github.com/mateo-berri) in [#​35748](https://github.com/BerriAI/litellm/pull/35748) - perf(streaming): assemble streamed tool-call arguments in linear time by [@​mateo-berri](https://github.com/mateo-berri) in [#​35826](https://github.com/BerriAI/litellm/pull/35826) - fix(s3\_v2): sign S3 object URLs with S3SigV4Auth so encoded paths verify by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35726](https://github.com/BerriAI/litellm/pull/35726) - test(e2e): self-seed the ui suite's password-login users in global setup by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35863](https://github.com/BerriAI/litellm/pull/35863) - fix(claude-code): create-only skill registration with a PUT update route (LIT-4110) by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​31752](https://github.com/BerriAI/litellm/pull/31752) - fix(proxy): fix zguard httpcode when block input by [@​jwang-gif](https://github.com/jwang-gif) in [#​31948](https://github.com/BerriAI/litellm/pull/31948) - fix(lint): pick the merge-aware base so in-progress merges are not blamed for base drift by [@​mateo-berri](https://github.com/mateo-berri) in [#​35868](https://github.com/BerriAI/litellm/pull/35868) - chore: bump litellm-proxy-extras 0.4.82 -> 0.4.83 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35877](https://github.com/BerriAI/litellm/pull/35877) - feat(ui): add Test Routing to the auto router create form by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35859](https://github.com/BerriAI/litellm/pull/35859) - fix(ui): derive auto-router preset tests from the bundled preset JSON by [@​tin-berri](https://github.com/tin-berri) in [#​35882](https://github.com/BerriAI/litellm/pull/35882) - revert: "test(e2e): vendor API strategy coverage across endpoints" ([#​34649](https://github.com/BerriAI/litellm/issues/34649)) by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35881](https://github.com/BerriAI/litellm/pull/35881) - chore(deps): bump grpc and golang.org/x modules in the terraform provider by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35844](https://github.com/BerriAI/litellm/pull/35844) - test(e2e): skip view-backed global spend probes pending LIT-5211 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35875](https://github.com/BerriAI/litellm/pull/35875) - fix(lint): move the basedpyright heap flag into the type check gate by [@​mateo-berri](https://github.com/mateo-berri) in [#​35869](https://github.com/BerriAI/litellm/pull/35869) - chore(ci): promote internal staging to main by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35876](https://github.com/BerriAI/litellm/pull/35876) - feat(ui): add role capability gating, migrate Tool Policies route by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35812](https://github.com/BerriAI/litellm/pull/35812) - refactor(ui): inject the fetch client's base url instead of reading it at import by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35802](https://github.com/BerriAI/litellm/pull/35802) - chore: remove unused .flake8 config and flake8 dev dependency by [@​mateo-berri](https://github.com/mateo-berri) in [#​35888](https://github.com/BerriAI/litellm/pull/35888) - chore: stop advising pre-commit and bootstrap by [@​mateo-berri](https://github.com/mateo-berri) in [#​35884](https://github.com/BerriAI/litellm/pull/35884) - fix(auth): name enable\_jwt\_auth when a JWT-shaped key is rejected by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35831](https://github.com/BerriAI/litellm/pull/35831) - feat(auto-router): make reminder marker pair configurable by [@​akapur99](https://github.com/akapur99) in [#​35874](https://github.com/BerriAI/litellm/pull/35874) - fix(UI): update anthropic model presets by [@​tin-berri](https://github.com/tin-berri) in [#​35896](https://github.com/BerriAI/litellm/pull/35896) - fix(bootstrap): switch to the dashboard node floor via nvm or fnm by [@​mateo-berri](https://github.com/mateo-berri) in [#​35895](https://github.com/BerriAI/litellm/pull/35895) - perf(pre-commit): run python, dashboard, and gen-api checks concurrently by [@​mateo-berri](https://github.com/mateo-berri) in [#​35903](https://github.com/BerriAI/litellm/pull/35903) - feat(spend): derive a default auto-router savings baseline from the hardest tier by [@​tin-berri](https://github.com/tin-berri) in [#​35907](https://github.com/BerriAI/litellm/pull/35907) - fix(http\_handler): self-heal handler clients closed after cache eviction by [@​mateo-berri](https://github.com/mateo-berri) in [#​35862](https://github.com/BerriAI/litellm/pull/35862) - fix(cost\_tracking): keep OpenAI prompt cache token details through usage reassembly by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34812](https://github.com/BerriAI/litellm/pull/34812) - fix(cost): bill gpt-5.6 prompt cache reads at the cache read rate by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34957](https://github.com/BerriAI/litellm/pull/34957) - fix(batches): account for Responses API usage by [@​rimysore](https://github.com/rimysore) in [#​35367](https://github.com/BerriAI/litellm/pull/35367) - ci: retry Codecov uploads and stop failing jobs on OIDC token flakes by [@​mateo-berri](https://github.com/mateo-berri) in [#​35251](https://github.com/BerriAI/litellm/pull/35251) - feat(complexity\_router): let operators rename the four complexity tiers by [@​akapur99](https://github.com/akapur99) in [#​35893](https://github.com/BerriAI/litellm/pull/35893) - chore(lint): zero stale ruff and LIT headroom and strip inert type: ignore comments by [@​mateo-berri](https://github.com/mateo-berri) in [#​35928](https://github.com/BerriAI/litellm/pull/35928) - chore(lint): zero out seven more purely local basedpyright rules by [@​mateo-berri](https://github.com/mateo-berri) in [#​35927](https://github.com/BerriAI/litellm/pull/35927) - chore(ui): zero stale headroom on local dashboard eslint budgets by [@​mateo-berri](https://github.com/mateo-berri) in [#​35929](https://github.com/BerriAI/litellm/pull/35929) - fix(managed-files): skip rows without file objects by [@​rimysore](https://github.com/rimysore) in [#​35365](https://github.com/BerriAI/litellm/pull/35365) - fix(router): redact fallback tracebacks at the call site and cover the sync deferred stream by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35843](https://github.com/BerriAI/litellm/pull/35843) - fix(migrations): recover from an interrupted Prisma toolchain install by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35832](https://github.com/BerriAI/litellm/pull/35832) - fix(lint): bring basedpyright rule counts back under their budget limits by [@​mateo-berri](https://github.com/mateo-berri) in [#​35962](https://github.com/BerriAI/litellm/pull/35962) - chore(ui): don't zero out stale headroom except no-console by [@​mateo-berri](https://github.com/mateo-berri) in [#​35964](https://github.com/BerriAI/litellm/pull/35964) - fix(proxy): give proxy\_admin\_viewer read parity with proxy\_admin by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35851](https://github.com/BerriAI/litellm/pull/35851) - refactor(ui): address UI lint budget issues by refactoring UI by [@​tin-berri](https://github.com/tin-berri) in [#​35960](https://github.com/BerriAI/litellm/pull/35960) - fix(ci): make the env-key doc gate see get\_secret\_bool reads by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35833](https://github.com/BerriAI/litellm/pull/35833) - fix(caching): re-land evicted LLM client closing ([#​35492](https://github.com/BerriAI/litellm/issues/35492)) atop self-healing handlers by [@​mateo-berri](https://github.com/mateo-berri) in [#​35870](https://github.com/BerriAI/litellm/pull/35870) - fix(proxy): keep the connected DB client when a startup health check fails by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35837](https://github.com/BerriAI/litellm/pull/35837) - chore(lint): remove litellm/types from the ruff lint exclusion by [@​mateo-berri](https://github.com/mateo-berri) in [#​35926](https://github.com/BerriAI/litellm/pull/35926) - feat(sgr): make the gateway middleware the source of truth for successful requests by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35717](https://github.com/BerriAI/litellm/pull/35717) - feat(auto-router): let operators replace the LLM classifier's system prompt by [@​akapur99](https://github.com/akapur99) in [#​35855](https://github.com/BerriAI/litellm/pull/35855) - fix(docker): bake the pip image's prisma engines at a world-readable path by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35976](https://github.com/BerriAI/litellm/pull/35976) - fix(auth): return 403 from the OAuth2 enterprise gate by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35838](https://github.com/BerriAI/litellm/pull/35838) - fix(router): keep custom model\_info across a price data reload by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35491](https://github.com/BerriAI/litellm/pull/35491) - fix(proxy): resolve pass-through credentials live from router deployments by [@​mateo-berri](https://github.com/mateo-berri) in [#​35916](https://github.com/BerriAI/litellm/pull/35916) - fix(ci): fetch only head and merge-base in lint jobs instead of every branch by [@​mateo-berri](https://github.com/mateo-berri) in [#​35982](https://github.com/BerriAI/litellm/pull/35982) - fix(autorouter): match CJK keyword\_tier\_rules that regex word boundaries miss by [@​akapur99](https://github.com/akapur99) in [#​35984](https://github.com/BerriAI/litellm/pull/35984) - feat(spend): rebuild the auto-router benchmarks backend as a per-session rollup by [@​tin-berri](https://github.com/tin-berri) in [#​35910](https://github.com/BerriAI/litellm/pull/35910) - refactor(ui): replace hand-rolled query-param routing with nuqs by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35871](https://github.com/BerriAI/litellm/pull/35871) - fix(docker): bake the componentized prisma engines at /opt/prisma so any uid can start by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35989](https://github.com/BerriAI/litellm/pull/35989) - fix(migrations): keep the toolchain heal from raising on an unreadable nodeenv cache by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35986](https://github.com/BerriAI/litellm/pull/35986) - fix(bedrock): sign Bedrock managed-file S3 requests with S3SigV4Auth by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35983](https://github.com/BerriAI/litellm/pull/35983) - chore(typing): replace Any seams with real types across responses, proxy, and provider adapters by [@​mateo-berri](https://github.com/mateo-berri) in [#​35809](https://github.com/BerriAI/litellm/pull/35809) - fix(ai21): resolve the documented AI21\_API\_KEY instead of a misspelled name by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35985](https://github.com/BerriAI/litellm/pull/35985) - fix(docker): fail the image build when the generated prisma engine paths drift off /opt/prisma by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35979](https://github.com/BerriAI/litellm/pull/35979) - fix(jina\_ai): resolve the documented JINA\_API\_KEY as a fallback by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35992](https://github.com/BerriAI/litellm/pull/35992) - fix(proxy): only treat a recoverable database outage as grounds to serve without one by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35864](https://github.com/BerriAI/litellm/pull/35864) - fix(ci): make every remaining CI checkout shallow by [@​mateo-berri](https://github.com/mateo-berri) in [#​35997](https://github.com/BerriAI/litellm/pull/35997) - fix(auto-router): stop the embedding model's context window from failing long requests by [@​akapur99](https://github.com/akapur99) in [#​35956](https://github.com/BerriAI/litellm/pull/35956) - fix(ci): make the env-key doc gate see bare get\_secret and get\_secret\_str reads by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35996](https://github.com/BerriAI/litellm/pull/35996) - fix(logging): extend secret redaction to records litellm does not emit directly by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35977](https://github.com/BerriAI/litellm/pull/35977) - test(utils): pin the register\_model replay test to the recorded half by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35994](https://github.com/BerriAI/litellm/pull/35994) - fix(ci): run every helm test suite, not just the first one per file by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35993](https://github.com/BerriAI/litellm/pull/35993) - ci: fail the build when a test file or Dockerfile is invoked by no job by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35991](https://github.com/BerriAI/litellm/pull/35991) - fix(langfuse): stop a collected httpx handler from closing a shared client by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35981](https://github.com/BerriAI/litellm/pull/35981) - fix(bedrock): grant bedrock:CountTokens in OIDC session policy by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​33145](https://github.com/BerriAI/litellm/pull/33145) - feat(pre-commit): save full lint output to a per-worktree log file by [@​mateo-berri](https://github.com/mateo-berri) in [#​36004](https://github.com/BerriAI/litellm/pull/36004) - feat(ui): match auto-router preset models against deployments' underlying model IDs by [@​tin-berri](https://github.com/tin-berri) in [#​35972](https://github.com/BerriAI/litellm/pull/35972) - fix(core\_helpers): map generic 'error' finish\_reason to 'stop' by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​33972](https://github.com/BerriAI/litellm/pull/33972) - fix(proxy)!: apply request-parameter checks consistently across body, path and form inputs by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36011](https://github.com/BerriAI/litellm/pull/36011) - fix: rebuild models\_by\_provider in add\_known\_models so cost map reloads reach wildcard expansion by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36010](https://github.com/BerriAI/litellm/pull/36010) - feat(complexity\_router): report LLM classifier cost per request via routing\_decision and x-litellm-classifier-cost header by [@​tin-berri](https://github.com/tin-berri) in [#​36015](https://github.com/BerriAI/litellm/pull/36015) - fix(model-prices): correct replicate model key typo by [@​AkashNaickar](https://github.com/AkashNaickar) in [#​34800](https://github.com/BerriAI/litellm/pull/34800) - fix(proxy): register managed batch output files on terminal retrieve by [@​Souravrajvi0](https://github.com/Souravrajvi0) in [#​34092](https://github.com/BerriAI/litellm/pull/34092) - perf(pre-commit): fetch basedpyright base counts from CI artifacts by [@​mateo-berri](https://github.com/mateo-berri) in [#​35970](https://github.com/BerriAI/litellm/pull/35970) - fix(ui): sync projects list page index to ?page= so back and reload keep the page by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36003](https://github.com/BerriAI/litellm/pull/36003) - fix(ui): link project page keys to their virtual key detail by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36002](https://github.com/BerriAI/litellm/pull/36002) - refactor(ui): drop unreferenced locals from dashboard route components by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35819](https://github.com/BerriAI/litellm/pull/35819) - fix(ui): opening a project now pushes ?project= so back and deep links work by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36001](https://github.com/BerriAI/litellm/pull/36001) - refactor(ui): drop unreferenced locals from shared dashboard components by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35821](https://github.com/BerriAI/litellm/pull/35821) - refactor(ui): drop unreferenced locals from tests and narrow destructures by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36025](https://github.com/BerriAI/litellm/pull/36025) - fix(guardrails): allow litellm\_content\_filter to run on post\_mcp\_call by [@​mateo-berri](https://github.com/mateo-berri) in [#​35980](https://github.com/BerriAI/litellm/pull/35980) - fix(guardrails): scan /v1/messages tool traffic by [@​mateo-berri](https://github.com/mateo-berri) in [#​35999](https://github.com/BerriAI/litellm/pull/35999) - refactor(ui): drop dead locals and unused React state across the dashboard by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36026](https://github.com/BerriAI/litellm/pull/36026) - feat(ui): add the auto-router usage tab to cost optimization by [@​tin-berri](https://github.com/tin-berri) in [#​35995](https://github.com/BerriAI/litellm/pull/35995) - fix(managed\_files): derive unified output file ids deterministically so concurrent registrations converge by [@​mateo-berri](https://github.com/mateo-berri) in [#​36019](https://github.com/BerriAI/litellm/pull/36019) - fix(proxy): send keepalive pings on anthropic messages SSE streams during upstream silence by [@​mateo-berri](https://github.com/mateo-berri) in [#​36024](https://github.com/BerriAI/litellm/pull/36024) - fix(managed\_files): return unified ids from unscoped file listing by [@​mateo-berri](https://github.com/mateo-berri) in [#​36031](https://github.com/BerriAI/litellm/pull/36031) - fix(arize\_phoenix): lowercase OTLP/gRPC auth metadata key by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34883](https://github.com/BerriAI/litellm/pull/34883) - fix(auto-router): accept every reminder marker pair a harness emits by [@​tin-berri](https://github.com/tin-berri) in [#​36029](https://github.com/BerriAI/litellm/pull/36029) - fix(pricing): sync flex/priority tier keys to dated OpenAI snapshot variants by [@​mateo-berri](https://github.com/mateo-berri) in [#​35923](https://github.com/BerriAI/litellm/pull/35923) - fix(cost): bill reasoning tokens at the service tier output rate by [@​mateo-berri](https://github.com/mateo-berri) in [#​35925](https://github.com/BerriAI/litellm/pull/35925) - fix(proxy): include today's UTC bucket when a daily activity range ends at the caller's current day by [@​tin-berri](https://github.com/tin-berri) in [#​36051](https://github.com/BerriAI/litellm/pull/36051) - fix: expired-miss share over all measured turns + cost-optimization tab labels by [@​tin-berri](https://github.com/tin-berri) in [#​36037](https://github.com/BerriAI/litellm/pull/36037) - fix(router): include Bedrock batch/S3 fields and model in deployment credentials by [@​mpcusack-altos](https://github.com/mpcusack-altos) in [#​24548](https://github.com/BerriAI/litellm/pull/24548) - fix(batch): track cost for managed batches with no attributable key/u… by [@​elinacse](https://github.com/elinacse) in [#​35468](https://github.com/BerriAI/litellm/pull/35468) - feat(guardrails): add scan\_only\_tool\_results to scope unified guardrails to tool results by [@​mateo-berri](https://github.com/mateo-berri) in [#​36014](https://github.com/BerriAI/litellm/pull/36014) - fix(cost): stop token-pricing the placeholder input on file content calls by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35140](https://github.com/BerriAI/litellm/pull/35140) - fix(proxy): fetch background responses through the router in CheckResponsesCost by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35137](https://github.com/BerriAI/litellm/pull/35137) - fix(proxy): yaml store\_prompts\_in\_spend\_logs should take precedence over DB cached value by [@​Praveena-617](https://github.com/Praveena-617) in [#​35769](https://github.com/BerriAI/litellm/pull/35769) - fix(lint): measure the basedpyright budget gate in a gate-owned venv by [@​mateo-berri](https://github.com/mateo-berri) in [#​36050](https://github.com/BerriAI/litellm/pull/36050) - docs: cap all GitHub comments at 15-25 words, curb semicolon splices by [@​mateo-berri](https://github.com/mateo-berri) in [#​36059](https://github.com/BerriAI/litellm/pull/36059) - chore(lint): name MappingProxyType in the mutable-collection fix messages by [@​mateo-berri](https://github.com/mateo-berri) in [#​36072](https://github.com/BerriAI/litellm/pull/36072) - test: roll back runtime model registrations between tests by [@​mateo-berri](https://github.com/mateo-berri) in [#​36039](https://github.com/BerriAI/litellm/pull/36039) - refactor(types): cut 653 implicit and explicit Any diagnostics across 11 modules by [@​mateo-berri](https://github.com/mateo-berri) in [#​36054](https://github.com/BerriAI/litellm/pull/36054) - fix(proxy): stop resolving the UI session sentinel team on /search\_tools/list by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36061](https://github.com/BerriAI/litellm/pull/36061) - fix(batches): persist managed file ids for cancelled/failed/expired batches by [@​mateo-berri](https://github.com/mateo-berri) in [#​36048](https://github.com/BerriAI/litellm/pull/36048) - fix(batches): register managed output files on batch cancel by [@​mateo-berri](https://github.com/mateo-berri) in [#​36034](https://github.com/BerriAI/litellm/pull/36034) - fix(proxy): allow non-admins to reach /user/daily/activity/aggregated by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36062](https://github.com/BerriAI/litellm/pull/36062) - fix(anthropic): coerce explicit additionalProperties to false in output\_format schema by [@​dkindlund](https://github.com/dkindlund) in [#​35811](https://github.com/BerriAI/litellm/pull/35811) - fix(batches): prevent managed file fallbacks by [@​rimysore](https://github.com/rimysore) in [#​35371](https://github.com/BerriAI/litellm/pull/35371) - chore: ignore the mechanical lint and typing sweeps in git blame by [@​mateo-berri](https://github.com/mateo-berri) in [#​36076](https://github.com/BerriAI/litellm/pull/36076) - fix(proxy): warn at startup when max\_budget is set but no database is connected by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36041](https://github.com/BerriAI/litellm/pull/36041) - fix(proxy): promote caller metadata trace fields into litellm\_metadata by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35866](https://github.com/BerriAI/litellm/pull/35866) - feat(terraform): sync provider 0.3.0 from the mirror and cut 0.4.0 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36098](https://github.com/BerriAI/litellm/pull/36098) - fix(guardrails): honor configured timeout in Zscaler AI Guard by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36110](https://github.com/BerriAI/litellm/pull/36110) - fix(logging): fall back to litellm\_metadata when metadata is empty by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36105](https://github.com/BerriAI/litellm/pull/36105) - fix(proxy): re-assert the authenticated identity on passthrough requests by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36121](https://github.com/BerriAI/litellm/pull/36121) - chore: bump litellm-enterprise 0.1.53 -> 0.1.54, litellm-proxy-extras 0.4.83 -> 0.4.84 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36139](https://github.com/BerriAI/litellm/pull/36139) - fix(ui): match auto-router preset models against wildcard-expanded model groups by [@​tin-berri](https://github.com/tin-berri) in [#​36111](https://github.com/BerriAI/litellm/pull/36111) - test(router): assert the auto-router max\_input\_chars kwarg by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36109](https://github.com/BerriAI/litellm/pull/36109) - fix(ui): allow clearing a key's budget reset from the Edit Key form by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36140](https://github.com/BerriAI/litellm/pull/36140) - fix(managed\_files): skip unparseable rows when listing managed files by [@​mateo-berri](https://github.com/mateo-berri) in [#​36021](https://github.com/BerriAI/litellm/pull/36021) - fix(a2a): stop writing per-caller headers onto the shared cached httpx client by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35978](https://github.com/BerriAI/litellm/pull/35978) - build(deps): bump h2 to 4.4.1 and js-yaml to 4.3.1 by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36147](https://github.com/BerriAI/litellm/pull/36147) - chore: promote staging to main by [@​mateo-berri](https://github.com/mateo-berri) in [#​36057](https://github.com/BerriAI/litellm/pull/36057) - fix(azure\_sentinel): respect AZURE\_AUTHORITY\_HOST and derive the Azure Monitor audience per cloud by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36137](https://github.com/BerriAI/litellm/pull/36137) - fix(bedrock): pass SSE-KMS key through to the batch input-file S3 upload by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35148](https://github.com/BerriAI/litellm/pull/35148) - fix(anthropic adapter): stop indexing choices\[0] on choiceless streaming chunks by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35314](https://github.com/BerriAI/litellm/pull/35314) - fix(bedrock): normalize /v1/completions and /v1/responses batch records by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35675](https://github.com/BerriAI/litellm/pull/35675) - fix(proxy): return the real status code when a credential update is rejected by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36166](https://github.com/BerriAI/litellm/pull/36166) - fix(proxy): improve Headroom /v1/compress HTTP 404 diagnostics by [@​aayush598](https://github.com/aayush598) in [#​35952](https://github.com/BerriAI/litellm/pull/35952) - fix(proxy): invalidate cached project object on project update and delete by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36028](https://github.com/BerriAI/litellm/pull/36028) - feat(proxy): add apply\_user\_budget\_to\_team\_keys opt-in by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36102](https://github.com/BerriAI/litellm/pull/36102) - fix(proxy): stop alerting on health probes that lose the planned engine-restart race by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​36141](https://github.com/BerriAI/litellm/pull/36141) - test(docker): gate the componentized gateway and backend images on an arbitrary-uid offline boot by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​36136](https://github.com/BerriAI/litellm/pull/36136) - fix(http): stop pooled clients persisting cookies on the aiohttp jar too by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​36149](https://github.com/BerriAI/litellm/pull/36149) - fix(router): bound fallback-walk work and error-log volume by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​36148](https://github.com/BerriAI/litellm/pull/36148) - ci: wire credential\_endpoints tests into the proxy endpoints job by [@​cursor](https://github.com/cursor)\[bot] in [#​36187](https://github.com/BerriAI/litellm/pull/36187) - docs(keys): document /key/info fields and clarify budget\_reset\_at is the next reset by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36127](https://github.com/BerriAI/litellm/pull/36127) - fix(azure\_sentinel): add AZURE\_SENTINEL\_AUTHORITY\_HOST as a Sentinel scoped override by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36165](https://github.com/BerriAI/litellm/pull/36165) - docs(pr-template): add a User Flow section with authoring instructions by [@​mateo-berri](https://github.com/mateo-berri) in [#​36162](https://github.com/BerriAI/litellm/pull/36162) - fix(proxy): derive config agent ids from agent\_name so grants survive secret rotation by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36020](https://github.com/BerriAI/litellm/pull/36020) - chore(ui): regenerate schema.d.ts for the /key/info docstring update by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36210](https://github.com/BerriAI/litellm/pull/36210) - build(deps): bump gitpython to 3.1.58 to clear osv-scan on staging by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36212](https://github.com/BerriAI/litellm/pull/36212) - fix(proxy): deny agent access when key and team grants resolve to nothing by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36221](https://github.com/BerriAI/litellm/pull/36221) - build(deps): defer the second pypdf advisory until the 6.15.0 bump by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36218](https://github.com/BerriAI/litellm/pull/36218) - fix(a2a): align agent list annotation and test with the tuple return type by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36217](https://github.com/BerriAI/litellm/pull/36217) - ci: always run the UI API types sync check so it can be required by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36213](https://github.com/BerriAI/litellm/pull/36213) - build(deps): bump nanoid to 3.3.17 in the dashboard lockfile by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36227](https://github.com/BerriAI/litellm/pull/36227) - feat(ui): show user email or alias in usage data export by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36232](https://github.com/BerriAI/litellm/pull/36232) - feat(auto-router): track turns per complexity tier (LIT-5302) by [@​tin-berri](https://github.com/tin-berri) in [#​36209](https://github.com/BerriAI/litellm/pull/36209) - fix(websearch): restore snippet text in native web\_search\_tool\_result blocks (LIT-5315) by [@​tin-berri](https://github.com/tin-berri) in [#​36228](https://github.com/BerriAI/litellm/pull/36228) - fix(proxy): resolve entity access groups in the model listing endpoints by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36230](https://github.com/BerriAI/litellm/pull/36230) - fix(ui): let access groups be a team's only model source, with hover provenance by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​36234](https://github.com/BerriAI/litellm/pull/36234) - fix(managed\_files): return unified output file ids from GET /batches by [@​mateo-berri](https://github.com/mateo-berri) in [#​36049](https://github.com/BerriAI/litellm/pull/36049) - test(proxy): compare empty agent list to the tuple get\_agent\_list returns by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36225](https://github.com/BerriAI/litellm/pull/36225) - fix(otel): name the RPC system and upstream on MCP tool-call spans by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35857](https://github.com/BerriAI/litellm/pull/35857) - fix(guardrails): chunk oversized Bedrock ApplyGuardrail requests instead of failing by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​36119](https://github.com/BerriAI/litellm/pull/36119) - test(e2e): settle control-plane writes across every replica, not just one by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36247](https://github.com/BerriAI/litellm/pull/36247) - fix(responses): forward allowed\_openai\_params through the chat completions bridge by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35885](https://github.com/BerriAI/litellm/pull/35885) - test(proxy): assert the copy \_add\_team\_member\_budget\_table returns by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36244](https://github.com/BerriAI/litellm/pull/36244) - chore(ui): regenerate dashboard api types for tier\_turns by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36243](https://github.com/BerriAI/litellm/pull/36243) - refactor(types): declare mirrored pricing fields on ModelInfo by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36215](https://github.com/BerriAI/litellm/pull/36215) - fix(lint): make strict-gate noqas survive base ruff and flag stale ones by [@​mateo-berri](https://github.com/mateo-berri) in [#​36257](https://github.com/BerriAI/litellm/pull/36257) - fix(vertex\_ai): surface real error/status on vertex batch create instead of IndexError 500 by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35141](https://github.com/BerriAI/litellm/pull/35141) - ci: give the remaining pull\_request workflows a concurrency group by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36252](https://github.com/BerriAI/litellm/pull/36252) - refactor(lint): graduate zero-violation strict rules and guard the budget ratchet by [@​mateo-berri](https://github.com/mateo-berri) in [#​36161](https://github.com/BerriAI/litellm/pull/36161) - fix(proxy): enforce require\_managed\_files on every route that accepts a raw provider id by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35551](https://github.com/BerriAI/litellm/pull/35551) - chore(typing): clear 1.4k basedpyright Any errors across 21 hotspot files by [@​mateo-berri](https://github.com/mateo-berri) in [#​36282](https://github.com/BerriAI/litellm/pull/36282) - test: roll back live router replay membership between tests by [@​mateo-berri](https://github.com/mateo-berri) in [#​36278](https://github.com/BerriAI/litellm/pull/36278) - chore(ci): sync main into internal staging by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36288](https://github.com/BerriAI/litellm/pull/36288) - build(lint): rename make pre-commit to make check with a working-tree fallback by [@​mateo-berri](https://github.com/mateo-berri) in [#​36277](https://github.com/BerriAI/litellm/pull/36277) - fix(ui): show team BYOK models in team fallback settings by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36241](https://github.com/BerriAI/litellm/pull/36241) - fix(otel): mark v2 server spans as failed for pre-call errors by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34546](https://github.com/BerriAI/litellm/pull/34546) - fix(websearch\_interception): bill intercepted searches to the calling key by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35708](https://github.com/BerriAI/litellm/pull/35708) - chore: remove pre-commit rule by [@​mateo-berri](https://github.com/mateo-berri) in [#​36295](https://github.com/BerriAI/litellm/pull/36295) - docs: clarify guideline priority ordering in CLAUDE.md by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​36296](https://github.com/BerriAI/litellm/pull/36296) - feat(router): independent, default-on deployment affinity for the auto-router by [@​tin-berri](https://github.com/tin-berri) in [#​36146](https://github.com/BerriAI/litellm/pull/36146) - test: repair stale CircleCI contracts by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36293](https://github.com/BerriAI/litellm/pull/36293) - chore(ci): promote internal staging to main by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36286](https://github.com/BerriAI/litellm/pull/36286) - chore: rebuild Admin UI bundle for the 2026-08-08 release by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36297](https://github.com/BerriAI/litellm/pull/36297) - chore(ci): promote internal staging to main by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​36304](https://github.com/BerriAI/litellm/pull/36304) ##### New Contributors - [@​rimysore](https://github.com/rimysore) made their first contribution in [#​35367](https://github.com/BerriAI/litellm/pull/35367) - [@​AkashNaickar](https://github.com/AkashNaickar) made their first contribution in [#​34800](https://github.com/BerriAI/litellm/pull/34800) - [@​Souravrajvi0](https://github.com/Souravrajvi0) made their first contribution in [#​34092](https://github.com/BerriAI/litellm/pull/34092) - [@​elinacse](https://github.com/elinacse) made their first contribution in [#​35468](https://github.com/BerriAI/litellm/pull/35468) - [@​aayush598](https://github.com/aayush598) made their first contribution in [#​35952](https://github.com/BerriAI/litellm/pull/35952) - [@​cursor](https://github.com/cursor)\[bot] made their first contribution in [#​36187](https://github.com/BerriAI/litellm/pull/36187) **Full Changelog**: <https://github.com/BerriAI/litellm/compare/v1.96.0...v1.97.0> ### [`v1.97.0`](https://github.com/BerriAI/litellm/releases/tag/v1.97.0) [Compare Source](https://github.com/BerriAI/litellm/compare/v1.96.2...v1.97.0) ##### Verify Docker Image Signature All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0). **Verify using the pinned commit hash (recommended):** A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key: ```bash cosign verify \ --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \ ghcr.io/berriai/litellm:v1.97.0 ``` **Verify using the release tag (convenience):** Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules: ```bash cosign verify \ --key https://raw.githubusercontent.com/BerriAI/litellm/v1.97.0/cosign.pub \ ghcr.io/berriai/litellm:v1.97.0 ``` Expected output: ``` The following checks were performed on each of these signatures: - The cosign claims were validated - The signatures were verified against the specified public key ``` *** ##### What's Changed - feat(proxy): resolve Cursor thinking/fast model-name suffixes on /cursor/chat/completions by [@​mateo-berri](https://github.com/mateo-berri) in [#​35554](https://github.com/BerriAI/litellm/pull/35554) - fix(team-callbacks): actually stop logging when disable\_logging is called by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35520](https://github.com/BerriAI/litellm/pull/35520) - refactor(lint): drop redundant !s f-string conversion flags and fix displaced import-group comments by [@​mateo-berri](https://github.com/mateo-berri) in [#​35546](https://github.com/BerriAI/litellm/pull/35546) - fix(proxy): backfill null user\_email on existing users during JWT auth by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34588](https://github.com/BerriAI/litellm/pull/34588) - feat(playground): add non-streaming response toggle by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35560](https://github.com/BerriAI/litellm/pull/35560) - feat(teams): apply default organization to new teams from default team settings by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35540](https://github.com/BerriAI/litellm/pull/35540) - fix(ui): block Playground page for viewer roles on direct URL access by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35676](https://github.com/BerriAI/litellm/pull/35676) - fix(caching): close evicted LLM clients so their connections are reclaimed by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35492](https://github.com/BerriAI/litellm/pull/35492) - chore(deps): update brace-expansion, postcss, and gitpython to current patch releases by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35692](https://github.com/BerriAI/litellm/pull/35692) - refactor(ui): rename the create MCP server component to PascalCase by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35686](https://github.com/BerriAI/litellm/pull/35686) - fix(openai): drop undefined Union from owns\_wrapped\_http\_client annotation by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35706](https://github.com/BerriAI/litellm/pull/35706) - fix(openai): drop the undefined Union from owns\_wrapped\_http\_client by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35704](https://github.com/BerriAI/litellm/pull/35704) - chore(ui): note Google's Agent Platform rename in vector store setup by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​28076](https://github.com/BerriAI/litellm/pull/28076) - fix(proxy): apply key/team router\_settings.model\_group\_alias by [@​yassin-berriai](https://github.com/yassin-berriai) in [#​35486](https://github.com/BerriAI/litellm/pull/35486) - feat(complexity\_router): default session affinity off and expose it in the UI by [@​tin-berri](https://github.com/tin-berri) in [#​35714](https://github.com/BerriAI/litellm/pull/35714) - fix(datadog): read team callback dd\_\* params from kwargs instead of blocked dynamic params ([#​35115](https://github.com/BerriAI/litellm/issues/35115) port) by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35687](https://github.com/BerriAI/litellm/pull/35687) - refactor(ui): extract the MCP create form's logic and field groups by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35694](https://github.com/BerriAI/litellm/pull/35694) - test(ui): tier the MCP create tests into unit and integration by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35697](https://github.com/BerriAI/litellm/pull/35697) - fix(proxy): redact credential headers from request logging copies by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35678](https://github.com/BerriAI/litellm/pull/35678) - feat(guardrails/rubrik): prompt moderation, response-text blocking, streaming buffer, failure logging by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35722](https://github.com/BerriAI/litellm/pull/35722) - fix(ui): render Responses API request and response in the logs drawer by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35718](https://github.com/BerriAI/litellm/pull/35718) - fix(ui): hide guardrail review buttons from non-admin users by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​27535](https://github.com/BerriAI/litellm/pull/27535) - feat(team): custom metadata validation hook for team create and update by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​33353](https://github.com/BerriAI/litellm/pull/33353) - ci(circleci): install a pinned Rust toolchain on the Linux jobs by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35519](https://github.com/BerriAI/litellm/pull/35519) - fix(bedrock): stop forwarding no-op toolSpec.strict to Converse by [@​tin-berri](https://github.com/tin-berri) in [#​35688](https://github.com/BerriAI/litellm/pull/35688) - fix(ui): reject an auto-router keyword rule left empty instead of dropping it by [@​tin-berri](https://github.com/tin-berri) in [#​35705](https://github.com/BerriAI/litellm/pull/35705) - fix(guardrails/rubrik): attribute blocked requests to the caller that made them by [@​yucheng-berri](https://github.com/yucheng-berri) in [#​35734](https://github.com/BerriAI/litellm/pull/35734) - fix(responses): forward client headers to the provider on /v1/responses by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34531](https://github.com/BerriAI/litellm/pull/34531) - feat(spend): add net auto-router savings to the cost-optimization dashboard by [@​tin-berri](https://github.com/tin-berri) in [#​35521](https://github.com/BerriAI/litellm/pull/35521) - chore(typing): clear basedpyright Any errors in budget reset, access groups, and cache settings by [@​mateo-berri](https://github.com/mateo-berri) in [#​35719](https://github.com/BerriAI/litellm/pull/35719) - fix(spend): read what a request cost from the record instead of pricing it again by [@​tin-berri](https://github.com/tin-berri) in [#​35736](https://github.com/BerriAI/litellm/pull/35736) - perf: install hiredis so redis-py parses replies with its C parser by [@​Classic298](https://github.com/Classic298) in [#​35709](https://github.com/BerriAI/litellm/pull/35709) - feat(ui): show auto-router savings on the cost-optimization dashboard by [@​tin-berri](https://github.com/tin-berri) in [#​35522](https://github.com/BerriAI/litellm/pull/35522) - perf: build log messages lazily so filtered-out log records cost nothing by [@​Classic298](https://github.com/Classic298) in [#​35703](https://github.com/BerriAI/litellm/pull/35703) - fix(proxy): retry model cost map fetch with Retry-After-aware backoff and keep current map on reload failure by [@​ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#​35739](https://github.com/BerriAI/litellm/pull/35739) - feat(otel): stamp service tier attributes on inference spans by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​35679](https://github.com/BerriAI/litellm/pull/35679) - fix(proxy): log the model cost map reload failure lazily by [@​tin-berri](https://github.com/tin-berri) in [#​35750](https://github.com/BerriAI/litellm/pull/35750) - fix(groq): translate web\_search\_options to the browser\_search tool by [@​hMED22](https://github.com/hMED22) in [#​34971](https://github.com/BerriAI/litellm/pull/34971) - feat(ui): add admin-configurable user banner by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35729](https://github.com/BerriAI/litellm/pull/35729) - fix(e2e): make spend-counter redis connection env-driven for non-cluster deployments by [@​yuneng-berri](https://github.com/yuneng-berri) in [#​35732](https://github.com/BerriAI/litellm/pull/35732) - fix(proxy): make /cursor/chat/completions work with Cursor agent mode by [@​tin-berri](https://github.com/tin-berri) in [#​34029](https://github.com/BerriAI/litellm/pull/34029) - fix(proxy): propagate user\_email and bind api\_key on JWT auth attribution paths by [@​devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#​34331](https://github.com/BerriAI/litellm/pull/34331) - chore(build): move the Admin UI toolchain to Node 24 by [@​yuneng-berri](https://g…
TLDR
Problem this solves:
How it solves it:
Relevant issues
serializeKeywordTierRulesstops discarding empty rules, so an unfilled row can never pass for a saved onecomplexity_router_configcontent, closing the gapauto_router_model_namingdocuments as its own purposeLinear ticket
Resolves LIT-5133
Pre-Submission checklist
Please complete all items before asking a LiteLLM maintainer to review your PR
@greptileaito re-request a review after pushing changes)Screenshots / Proof of Fix
Live proxy on localhost:4000 against a real Postgres. First, what the two halves looked like before
That 500 leaves a durable row the router will never load, which is what the client-side strip was avoiding by throwing the rule away instead
After the change, on a fresh database. The same payload is refused at the boundary
Whitespace-only keywords hit the router's own normalizer, and a valid rule still creates and serves
Every write shape, against a fresh database. A patch that carries only a config is covered too, which is what the dashboard's edit modal and any caller updating just the routing rules send
Case 4 is the one worth calling out: before it was closed, that patch overwrote a serving router with a config it could not build and took it out of service on the way to the 500. Case 7 is why a stored config is never judged on a write that does not carry one
The inference leg is not included because both shared provider accounts are out of credit right now; routing behaviour is untouched by this PR either way
UI runbook
Start the proxy and
npm run devinui/litellm-dashboard, theninvoiceinto Keywords 1 and press enter. The error clears and the button comes back; click it and the rule is on the saved configType
🐛 Bug Fix
Changes
getKeywordTierRulesErrorrejects any rule whose keywords are empty after trimming and names each offending row;getSemanticConfigErrordrops the per-rule check it now subsumes, which only ever ran with the semantic toggle onserializeKeywordTierRulesreturns one entry per rule instead of filtering empties out, so validation and the payload agree on what a keyword is and row numbers line up with the "Keywords N" labelsadd_auto_router_tab.tsxandedit_auto_router_modal.tsxcall the new check before building their payloadKeywordTierRulesmarks an offending row withstatus="error"and "At least one keyword is required", driven by the sameemptyKeywordTierRuleIndexesthe submit message uses, so the row named and the row marked cannot divergeopen={false}leaves antd nothing for Enter to select, so the word previously only became a tag on blur, and clicking submit was what supplied that blur; with the button withheld that route is gone and the row could not be filledshowValidationErrorsstate it was missing, so its classifier error renders at allvalidate_complexity_router_config_writeparses a writtencomplexity_router_configwithComplexityRouterConfig, so/model/new,/model/{id}/updateand the legacy/model/updateall reject an unloadable one with a 400 and persist nothingThings a reviewer will ask about
Whether parsing at the write boundary can reject a config that works today. It cannot:
resolve_complexity_router_pluginsruns only on the config.yaml path, so a written config already reachesComplexityRouterConfig.model_validateunchanged moments later at load time. The boundary calls that same model rather than a copy of its rules, and the model isextra="allow", so unknown keys stay the caller's businessWhy the row reports itself instead of waiting for a failed submit. With the button withheld there is no failed submit left to surface anything, so a row that only spoke up after one would never speak at all. A row exists only because the caller clicked "Add keyword rule", so it is safe to have it report straight away; tiers keep their existing behaviour and rely on the tooltip, since a blank form turning red on load helps nobody
Why the serializer no longer drops empty rules. Dropping them is what made the bug invisible, and it breaks the row numbering the message depends on. Reverting the serializer alone fails five of the new tests, including the validator ones, which is the check that the call into it is real
Why the config is judged on its own rather than by classifying the deployment's model. A patch may write a config without naming a model, and the stored model is encrypted at rest, so it cannot be classified from the row. Judging the config alone is the only reading that covers that path
Why a write that carries no config is never rejected for a stored one. A row stored before this validation existed is already unloadable, and blocking its rename would break the restore path
_strategy_router_write_violationdocuments. The repair is a write that supplies a good config, which is coveredQA runbook
Follow the UI runbook above for the dashboard half, and the curl block for the API half
npx vitest run src/components/add_model src/components/edit_auto_routeris green; the full dashboard suite passes at 5906 testspytest tests/test_litellm/router_utils/test_auto_router_model_naming.py tests/test_litellm/proxy/management_endpoints/test_model_management_endpoints.py tests/test_litellm/router_strategy/test_complexity_router.pypasses at 397Mutation checks run against this diff:
Final Attestation
Note
Medium Risk
Touches proxy model create/update validation for auto-router deployments; behavior is aligned with existing router load rules and mainly prevents bad writes, but incorrect validation could block legitimate configs.
Overview
Fixes LIT-5133: empty keyword tier rules could be saved (or stripped client-side) and leave deployments that fail at router reload with a 500 instead of a clear validation error.
Proxy: Model create/update now runs
validate_complexity_router_config_writeon any incomingcomplexity_router_config, using the sameComplexityRouterConfigthe router loads. Invalid configs (e.g.keyword_tier_ruleswith no keywords) return 400 before persistence._strategy_router_write_violationvalidates config-only patches without requiringlitellm_params.model, and does not judge stored config on renames that omit config.Dashboard:
getKeywordTierRulesErrorandemptyKeywordTierRuleIndexesblock submit whenever a keyword rule has no non-empty keyword, including when semantic matching is off.serializeKeywordTierRuleskeeps empty rules in the payload instead of dropping them.KeywordTierRulesshows per-row errors, disables submit via sharedsubmitBlockedReason(tooltip on create/edit), and commits typed keywords on Enter/blur so rows can be filled with the button disabled. Create and edit flows both call the new checks before building the config.Reviewed by Cursor Bugbot for commit b647947. Bugbot is set up for automated code reviews on this repo. Configure here.