Skip to content

feat(team): custom metadata validation hook for team create and update - #33353

Merged
yuneng-berri merged 16 commits into
litellm_internal_stagingfrom
litellm_/bold-mclaren-8d89b3
Aug 4, 2026
Merged

feat(team): custom metadata validation hook for team create and update#33353
yuneng-berri merged 16 commits into
litellm_internal_stagingfrom
litellm_/bold-mclaren-8d89b3

Conversation

@yuneng-berri

@yuneng-berri yuneng-berri commented Jul 15, 2026

Copy link
Copy Markdown
Contributor

Relevant issues

Linear ticket

Resolves LIT-3979

Pre-Submission checklist

Please complete all items before asking a LiteLLM maintainer to review your PR

  • I have added meaningful tests
  • My PR passes all CI/CD checks (e.g., lint, format, unit tests)
  • My PR's scope is as isolated as possible; it only solves 1 specific problem
  • I have received a Greptile Confidence Score of at least 4/5 before requesting a maintainer review (Greptile reviews automatically once the PR is opened; only comment @greptileai to re-request a review after pushing changes)

Screenshots / Proof of Fix

Captured at commit b67cdb6 against a live proxy (port 4010) backed by Postgres, no mocks. The proxy was started with this config:

general_settings:
  master_key: sk-1234
  custom_team_metadata_validate: qa_team_metadata_validator.validate_team_metadata
  team_metadata_validation_timeout: 5
  team_metadata_validation_error_message: "Cost center validation is unavailable right now; the team was not saved. Contact FinOps."

and this validator next to it, which accepts cost centers CC-1001/CC-1002, rejects everything else with a specific message, and raises on CC-KABOOM to simulate the operator's upstream service being down:

from litellm.proxy.management_helpers.team_metadata_validation import (
    TeamMetadataValidationPayload,
    TeamMetadataValidationResult,
)

VALID_COST_CENTERS = {"CC-1001", "CC-1002"}

async def validate_team_metadata(payload: TeamMetadataValidationPayload) -> TeamMetadataValidationResult:
    cost_center = payload.metadata.get("cost_center")
    if cost_center == "CC-KABOOM":
        raise RuntimeError("simulated internal validation service outage")
    if cost_center is None:
        return TeamMetadataValidationResult(valid=False, error_message="Team metadata must include a cost_center. Contact the FinOps team.")
    if cost_center not in VALID_COST_CENTERS:
        return TeamMetadataValidationResult(valid=False, error_message=f"Cost center {cost_center} is not recognized. Contact the FinOps team.")
    return TeamMetadataValidationResult(valid=True)
  1. Create with a valid cost center succeeds
$ curl -sS -X POST 'http://localhost:4010/team/new' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_alias": "qa-valid-team", "team_id": "qa-team-valid-1", "metadata": {"cost_center": "CC-1001"}}'
{"team_alias":"qa-valid-team","team_id":"qa-team-valid-1",...,"metadata":{"cost_center":"CC-1001"},...}
  1. Create with an unrecognized cost center is blocked with the validator's own message
$ curl -sS -X POST 'http://localhost:4010/team/new' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_alias": "qa-bad-team", "metadata": {"cost_center": "CC-9999"}}'
{"error":{"message":"{'error': 'Cost center CC-9999 is not recognized. Contact the FinOps team.'}","type":"internal_server_error","param":"None","code":"400"}}
  1. Create with no metadata at all still validates, so a required key can be enforced
$ curl -sS -X POST 'http://localhost:4010/team/new' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_alias": "qa-no-meta-team"}'
{"error":{"message":"{'error': 'Team metadata must include a cost_center. Contact the FinOps team.'}","type":"internal_server_error","param":"None","code":"400"}}
  1. A raising validator fails closed with the configured generic message as a 503
$ curl -sS -X POST 'http://localhost:4010/team/new' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_alias": "qa-outage-team", "metadata": {"cost_center": "CC-KABOOM"}}'
{"error":{"message":"{'error': 'Cost center validation is unavailable right now; the team was not saved. Contact FinOps.'}","type":"internal_server_error","param":"None","code":"503"}}
  1. PATCH of an unrelated metadata key validates the merged result, so the preserved cost center passes
$ curl -sS -X PATCH 'http://localhost:4010/team/qa-team-valid-1' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"metadata": {"team_notes": "hello"}}'
{"team_alias":"qa-valid-team","team_id":"qa-team-valid-1",...,"metadata":{"team_notes":"hello","cost_center":"CC-1001"},...}
  1. Deleting the required key via PATCH null is caught, because the validator sees the key's absence in the merged result
$ curl -sS -X PATCH 'http://localhost:4010/team/qa-team-valid-1' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"metadata": {"cost_center": null}}'
{"error":{"message":"{'error': 'Team metadata must include a cost_center. Contact the FinOps team.'}","type":"internal_server_error","param":"None","code":"400"}}
  1. POST /team/update replacing metadata wholesale without the required key is caught the same way
$ curl -sS -X POST 'http://localhost:4010/team/update' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_id": "qa-team-valid-1", "metadata": {"team_notes": "only-notes"}}'
{"error":{"message":"{'error': 'Team metadata must include a cost_center. Contact the FinOps team.'}","type":"internal_server_error","param":"None","code":"400"}}
  1. An update that does not touch metadata skips validation and succeeds
$ curl -sS -X POST 'http://localhost:4010/team/update' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_id": "qa-team-valid-1", "tpm_limit": 50}'
{"team_id":"qa-team-valid-1","data":{...}}
  1. A rejected create leaves no row behind
$ curl -sS -X POST 'http://localhost:4010/team/new' ... -d '{"team_alias": "qa-reject-rowcheck", "team_id": "qa-team-rejected-1", "metadata": {"cost_center": "CC-9999"}}'   # rejected with 400
$ curl -sS 'http://localhost:4010/team/info?team_id=qa-team-rejected-1' -H 'Authorization: Bearer sk-1234'
{"error":{"message":"{'message': 'Team not found, passed team id: qa-team-rejected-1.'}","type":"auth_error","param":"None","code":"404"}}
  1. The blocked deletes in 6 and 7 left the stored metadata intact
$ curl -sS 'http://localhost:4010/team/info?team_id=qa-team-valid-1' -H 'Authorization: Bearer sk-1234' | python -c "import json,sys; print(json.load(sys.stdin)['team_info']['metadata'])"
{'team_notes': 'hello', 'cost_center': 'CC-1001'}

Review follow-ups (commit 93d802c)

Both bot findings were confirmed and fixed. Validation now runs before the model_aliases table insert, so a rejected create can no longer leave an orphaned LiteLLM_ModelTable row (regression test asserts the model create mock is never awaited on rejection). The payload's existing_metadata now has system-managed keys stripped from a copy, matching the stripped metadata field, so a key-preservation validator never sees server-owned keys like team_member_budget_id "disappear" on PATCH. The remaining Greptile note about a validator's own HTTPException being converted to the generic 503 describes the intended fail-closed contract: validators signal rejections by returning valid=False with their message, and any raise is treated as a system failure

The runner also accepts class instances exposing an async call (the natural shape for a validator holding an HTTP client), which previously failed the coroutine check

This commit adds a validator implementation matrix: three independent implementations (a static allowlist function, an HTTP-service-backed function using httpx against a stub cost center service, and an immutability-enforcing class instance that uses existing_metadata) each driven through the real /team/new, POST /team/update, and PATCH /team/{team_id} paths across 8 scenarios, plus service-outage coverage for the HTTP implementation. The same three implementations were also exercised against a live DB-backed proxy via config, with results matching the test matrix cell for cell

DB-backed proxy e2e in CI (commit f240a58)

The validator matrix now also runs full e2e against a Postgres-backed proxy in the existing proxy_store_model_in_db_tests CircleCI job. The job's config (store_model_db_config.yaml) registers team_metadata_validator_e2e.validate_team_metadata, a dispatching validator mounted next to the config in the container; it routes each request to one of the three implementations via a _e2e_validator_impl metadata key and accepts any request that does not carry the key, so the suite's other team operations are unaffected. CI starts a stand-in cost center service on the host (tests/store_model_in_db_tests/cost_center_service.py, port 9414) which the HTTP-backed implementation reaches through host.docker.internal, mirroring how the job already runs its fake OpenAI endpoint; the outage scenario targets a closed port so the fail-closed 503 is proven without stopping services. The job already passes LITELLM_LICENSE, which this premium-gated feature requires

tests/store_model_in_db_tests/test_team_metadata_validation_e2e.py holds the matrix: 8 scenarios x 3 implementations plus the outage case and a no-dispatch-key control, asserting status codes, the validator-authored messages, that a rejected create leaves no team row, and that a blocked update leaves stored metadata intact. The full file was verified locally against a DB-backed proxy started with the same config and mounted validator (26/26 passed) before wiring it into CI

Admin UI: metadata as key-value pairs (commit eda8491)

The team create and edit forms previously asked for metadata as raw JSON in a textarea buried inside Additional Settings, so a malformed blob failed in the browser before the validator ever saw it and admins had to hand-write JSON for what is conceptually a set of key-value pairs. Both forms now render a key-value pair editor directly under the TPM/RPM limit fields, backed by a new shared MetadataKeyValueFields component. Values accept plain text or JSON: non-string values display as JSON and parse back to their typed form on save, and JSON-ambiguous strings are quoted on display so types survive the round trip. The edit form hides UI-managed keys (logging, guardrails, model rate limits and similar) that dedicated controls already own and re-add on save

Verified end to end at eda8491 with the dashboard dev server against a live Dockerized proxy on localhost:4000 running this PR's allowlist cost center validator (accepts CC-1001/CC-1002):

  1. Teams -> Create Team: fill Team Name and Models, add the pair cost_center / CC-9999 in the Metadata editor under the RPM field, submit; the create is blocked and the toast carries the validator's own message: Error creating the team: ApiError: {'error': 'Cost center CC-9999 is not recognized. Contact the FinOps team.'}
  2. Change the value to CC-1001 and resubmit; the team is created and GET /team/list shows metadata {"cost_center": "CC-1001"} stored as a typed object
  3. Open the team, Settings tab, Edit Settings: the pair prefills as cost_center / CC-1001 under the RPM field, and UI-managed keys do not appear as editable pairs
  4. Change the value to CC-9999 and Save Changes; the update is blocked with the same validator message and the form stays open for correction
  5. Change the value to CC-1002 and save; GET /team/info confirms cost_center is now CC-1002 while the UI-managed keys are untouched

Before/after screenshots (before: the JSON textarea inside Additional Settings at f240a58; after: the pair editor under the RPM field at eda8491) reproduce from http://localhost:3000/teams via Create Team for the create modal and via any team's Settings -> Edit Settings for the edit form, with the proxy from the QA config above on localhost:4000

Schema-prepopulated metadata fields (commits 19c2577 through eea99b3)

Customer feedback on the dev image was that users filling in team metadata have no indication of which keys are expected or how they are named (cost_center vs CostCenter vs costcenter). These commits add a declarative team_metadata_schema in config that the UI uses to prepopulate the metadata editor, and clean up how a validator rejection is shown to the user. Each schema field is a key plus an optional display label. Verified at 698bc8a against a live DB-backed proxy running both the schema and this PR's cost center validator:

general_settings:
  master_key: sk-1234
  custom_team_metadata_validate: qa_team_metadata_validator.validate_team_metadata
  team_metadata_validation_timeout: 5
  team_metadata_validation_error_message: "Cost center validation is unavailable right now; the team was not saved. Contact FinOps."
  team_metadata_schema:
    - key: cost_center
      label: "Cost Center"
    - key: app_name
      label: "Application Name"
    - key: chargeback_mode
      label: "Chargeback Mode"
  1. The new endpoint serves the declared fields to authenticated callers
$ curl -sS 'http://localhost:4000/team/metadata_schema' -H 'Authorization: Bearer sk-1234'
{"fields":[{"key":"cost_center","label":"Cost Center"},{"key":"app_name","label":"Application Name"},{"key":"chargeback_mode","label":"Chargeback Mode"}]}
  1. The endpoint requires auth; the schema is org-internal and deliberately not on the unauthenticated well-known UI config
$ curl -sS -o /dev/null -w "%{http_code}\n" http://localhost:4000/team/metadata_schema
401
$ curl -sS -o /dev/null -w "%{http_code}\n" http://localhost:4000/team/metadata_schema -H 'Authorization: Bearer sk-wrong'
401
  1. The schema and the validator compose: the same proxy rejects a create whose cost center the validator does not know, with the validator's own message and no row written
$ curl -sS -X POST 'http://localhost:4000/team/new' -H 'Authorization: Bearer sk-1234' -H 'Content-Type: application/json' \
    -d '{"team_alias": "qa-reject-check", "metadata": {"cost_center": "CC-9999"}}'
{"error":{"message":"{'error': 'Cost center CC-9999 is not recognized. Contact the FinOps team.'}","type":"internal_server_error","param":"None","code":"400"}}
  1. With no schema configured the endpoint returns an empty list, which the UI renders as the plain free-form editor, so nothing changes for proxies that do not opt in
$ curl -sS 'http://localhost:4042/team/metadata_schema' -H 'Authorization: Bearer sk-1234'
{"fields":[]}
  1. A malformed schema fails proxy startup instead of silently serving a broken form (unknown field names are rejected)
$ python litellm/proxy/proxy_cli.py --config bad_schema_config.yaml --port 4043; echo "exit code: $?"
pydantic_core._pydantic_core.ValidationError: 1 validation error for tuple[TeamMetadataFieldSchema, ...]
  Extra inputs are not permitted [type=extra_forbidden, input_value=True, input_type=bool]
exit code: 3

UI screenshots reproduce from http://localhost:3000/teams with the dev server pointed at a proxy running the config above:

  1. Click Create Team and scroll to Metadata: each declared key is prepopulated as an ordinary key-value pair row, exactly as if it had been added with the Add Key-Value Pair button, with the key filled in; rows stay fully editable and removable
  2. Enter CC-9999 as the cost center and submit: the server-side validator rejects it and the toast reads "Error creating the team: Cost center CC-9999 is not recognized. Contact the FinOps team." with no ApiError or JSON wrapper around the message
  3. Change to CC-1001 and create, then open the team, Settings, Edit Settings: cost_center shows as a prefilled pair row alongside any other stored keys, and declared keys missing from the stored metadata are appended as empty rows; a rejected save shows the same clean message via "Failed to update team settings: ..."
  4. Add another row and give it the key cost_center: the ordinary duplicate-key validation rejects it
  5. Fill Application Name and save, then blank it and save again: the key appears in stored metadata after the first save and is gone after the second, since blank fields are omitted rather than saved as empty strings
  6. Remove team_metadata_schema from the config, restart the proxy, and reopen either modal: the plain free-form editor is back, and any metadata keys the schema used to own reappear as ordinary editable rows

Type

🆕 New Feature

Changes

Operators can now require and validate custom team metadata (for example a cost center that must exist in an internal system of record) by pointing the proxy config at their own async Python function, following the same config surface as custom_key_generate:

general_settings:
  custom_team_metadata_validate: my_module.validate_team_metadata
  team_metadata_validation_timeout: 5              # optional, seconds, default 5
  team_metadata_validation_error_message: "..."    # optional, generic fail-closed message

The function receives a typed TeamMetadataValidationPayload (operation, the metadata that will actually be written, the stored metadata, team id/alias, and requester identity) and returns a TeamMetadataValidationResult. Returning valid=False blocks the write with the function's error_message as a 400. Any raised exception, timeout, or malformed return fails closed with the configurable generic message as a 503, so an unreachable upstream validation service can never let an unvalidated team through

The hook gates every team write path: POST /team/new (always, even when no metadata is sent, so a required key can be enforced at creation), POST /team/update, and PATCH /team/{team_id} (only when the request carries metadata; the validator sees the RFC 7386 merged result on PATCH and the replacement on POST). It runs before any side effects, so a rejected write leaves no rows behind. The loaded validator lives in a small registry object rather than a new module global, and the new config key is registered alongside the sibling custom hook keys in the DB-overlay config handling for consistent behavior

The feature is premium-gated, matching enforced_params. An example validator is included at litellm/proxy/example_config_yaml/custom_team_metadata_validate.py

New files: litellm/proxy/management_helpers/team_metadata_validation.py plus its mapped test file. Touched: proxy_server.py (config load), team_endpoints.py (two call sites), and the team endpoint tests

The Admin UI now edits team metadata as key-value pairs in both the team create and edit forms, placed directly under the TPM/RPM limit fields instead of a JSON textarea inside Additional Settings. A shared MetadataKeyValueFields component owns the pair rows and the object/pairs conversion, values may be plain text or JSON and round-trip losslessly, and the edit form filters out UI-managed metadata keys that dedicated controls re-add on save. Touched: Teams.tsx and team/TeamInfo.tsx plus their mapped test files, and the new component with its own test file

Building on the validator, operators can now declare the expected team metadata keys in general_settings.team_metadata_schema (each entry is a key plus an optional display label). A new authenticated GET /team/metadata_schema endpoint serves the declaration; the config is parsed with pydantic at load, so an unknown field name, a missing or empty key, or duplicate keys fail proxy startup. The schema is advisory and display-only; enforcement stays entirely with custom_team_metadata_validate

In both team modals, declared keys are prepopulated into the ordinary key-value editor as if the user had added the rows themselves: the key input is prefilled with the declared key and the rows stay fully editable and removable. Keys already present in the stored metadata are not duplicated, and deleting or renaming a prepopulated row behaves exactly like any hand-added row. The dashboard fetches the schema when the Teams page mounts (react-query, 24 hour staleTime matching the well-known UI config, retry 1, since config changes require a restart anyway) and shows a skeleton in the metadata section while the fetch is in flight. On any fetch error the modals silently fall back to the free-form editor with no toast or console noise: the server-side validator still fails closed, so a degraded UI can never let invalid metadata through, and a proxy without the endpoint or schema behaves exactly as before

The endpoint is registered in the info route group alongside /team/info and /team/available (commit 0d0b486), so team admins and internal users editing their team see the prepopulated keys too; without this the route classified as admin-only management and non-admin dashboard sessions got a 401, which the fail-open path silently turned into a plain free-form editor. Found by a live-key test matrix; pinned by a route-group membership test and a TestClient auth test

The validation module was also mutation tested (mutmut, 114 mutants): the five initial survivors exposed a missing timeout boundary case, untested effective-timeout wiring, and loose error-message assertions, and eea99b3 adds tests that bring the module to a 100 percent kill rate

Validator rejections now surface cleanly in the UI. The proxy serializes an HTTPException detail as the string form of a Python dict, so the browser previously toasted "Error creating the team: ApiError: {'error': ...}" on create and the whole raw JSON envelope on update. A new unwrapProxyErrorMessage helper in the shared http client unwraps both shapes down to the validator's own message, and the create and update toasts use it

New files: the TeamMetadataFieldSchema/TeamMetadataSchemaResponse types, parse_team_metadata_schema plus a schema registry in team_metadata_validation.py, the useTeamMetadataSchema react-query hook, and their mapped test files. Touched: proxy_server.py (config load), team_endpoints.py (the GET route), MetadataKeyValueFields.tsx (schema rows, skeleton, free-form guard), both team modals, the shared http client error helpers, and the regenerated dashboard API types

Final Attestation

  • The tests check the right things, including the edge cases, and regressions in the respective real-world customer use-cases are not possible after this PR

Operators can point general_settings.custom_team_metadata_validate at an
async Python function that validates team metadata before /team/new,
POST /team/update, and PATCH /team/{team_id} commit their writes. The
hook receives the metadata that will actually be written (the merged
result on PATCH) plus the stored metadata and requester context, and
fails closed: a rejected value returns the function's own message as a
400 while any exception or timeout blocks the write with a configurable
generic message as a 503. Premium-gated like enforced_params.
@greptile-apps

greptile-apps Bot commented Jul 15, 2026

Copy link
Copy Markdown
Contributor

Greptile Summary

This PR adds a custom_team_metadata_validate hook that gates every team write path (POST /team/new, POST /team/update, PATCH /team/{team_id}) with an operator-supplied async validator, following the same config surface and registry pattern as custom_key_generate. It also introduces a team_metadata_schema config that prepopulates the metadata editor in both team modals, and replaces the raw-JSON textarea in those modals with a key-value pair editor backed by a new MetadataKeyValueFields component.

  • Backend: Validation runs before any DB side effects (model table insert, team table create/update), with system-managed metadata keys stripped from both metadata and existing_metadata before the validator sees them. A rejected call returns HTTP 400 with the validator's own message; any raised exception or timeout fails closed as HTTP 503 with a configurable generic message. The premium gate, registry reset on config reload, and DB-overlay config registration all match the sibling hook patterns.
  • Frontend: The pair editor supports plain-text and JSON-typed values with lossless round-tripping, schema-declared keys are prepopulated as ordinary editable rows, and validator rejections are now surfaced as the validator's own message rather than the raw JSON envelope via unwrapProxyErrorMessage.
  • Tests: A 704-line unit test file with mutation-tested coverage, a validator implementation matrix tested against live endpoints, and a CircleCI e2e job covering DB-backed scenarios.

Confidence Score: 5/5

Safe to merge; the validation hook is additive and opt-in, all write paths fail closed, and existing teams without the config key are completely unaffected.

The core validation logic, registry/cleanup plumbing, and UI changes are all correct and well-tested. The only finding is that the HTTP-backed validator scenarios in the matrix test start a real ThreadingHTTPServer and make live httpx calls, which is a policy mismatch for the tests/test_litellm/ folder but does not affect production behavior or other tests.

Files Needing Attention: tests/test_litellm/proxy/management_helpers/test_team_metadata_validation.py — the matrix test scenarios for the HTTP implementation use real socket I/O and should be moved or mocked.

Important Files Changed

Filename Overview
litellm/proxy/management_helpers/team_metadata_validation.py New module implementing the custom team metadata validation hook; fail-closed design, registry pattern, schema parsing, and timeout/unavailable-message helpers all look correct
litellm/proxy/management_endpoints/team_endpoints.py Validation inserted at the right points (before any DB side effects on create; on update/patch when metadata is present); system-managed key stripping applied to both metadata and existing_metadata before the validator call
litellm/proxy/proxy_server.py Validator and schema registries are loaded and reset correctly alongside sibling hooks; new config key registered in DB-overlay remote module string fields for consistent behavior
tests/test_litellm/proxy/management_helpers/test_team_metadata_validation.py Comprehensive test coverage including a validator implementation matrix, but the matrix tests for the HTTP-backed implementation start a real ThreadingHTTPServer and make real httpx connections, violating the no-network-calls policy for this test folder
ui/litellm-dashboard/src/lib/http/client.ts unwrapProxyErrorMessage and extractProxyErrorMessage helpers correctly peel the Python-dict-string wrapper from proxy error responses; recursive JSON unwrap terminates because the derived-equals-trimmed check prevents re-parsing the same value
ui/litellm-dashboard/src/components/common_components/MetadataKeyValueFields.tsx New shared component for key-value pair metadata editing; schema prepopulation, JSON round-trip, and skeleton loading state all look correct

Reviews (5): Last reviewed commit: "fix(proxy): use pooled async httpx clien..." | Re-trigger Greptile

Comment thread litellm/proxy/management_helpers/team_metadata_validation.py
Comment thread litellm/proxy/management_endpoints/team_endpoints.py
Comment thread litellm/proxy/management_endpoints/team_endpoints.py Outdated
@veria-ai

veria-ai Bot commented Jul 15, 2026

Copy link
Copy Markdown
Contributor

PR overview

This PR adds a custom validation hook for team metadata during team creation and updates in the proxy management endpoints. It focuses on applying configured metadata rules as team records are created or modified.

There is one remaining issue: some metadata-backed team fields can still be supplied outside the explicit metadata object and persisted after validation runs, allowing a team admin to bypass the configured metadata policy. One issue has already been addressed, so the PR is moving in the right direction, but the remaining gap affects the core validation behavior this change introduces. The risk is limited by requiring team-admin-level action, but it can still result in policy-violating team metadata being saved.

Open issues (1)

Fixed/addressed: 1 · PR risk: 5/10

@codspeed-hq

codspeed-hq Bot commented Jul 15, 2026

Copy link
Copy Markdown
Contributor

Merging this PR will not alter performance

✅ 31 untouched benchmarks


Comparing litellm_/bold-mclaren-8d89b3 (bb645aa) with litellm_internal_staging (ba1bde7)1

Open in CodSpeed

Footnotes

  1. No successful run was found on litellm_internal_staging (c6a796a) during the generation of this report, so ba1bde7 was used instead as the comparison base. There might be some changes unrelated to this pull request in this report.

…em keys from validator input

Review follow-ups on the team metadata validation hook: run the validator
before the model_aliases table insert so a rejected create leaves no
orphaned model rows, strip system-managed keys from existing_metadata so
the validator sees symmetric input on both fields, and accept class
instances exposing an async __call__ as validators. Adds a three-way
validator implementation matrix (allowlist function, HTTP-service-backed
function, immutability-enforcing class instance) driven through the real
create, update, and patch endpoints, including an HTTP stub service and
outage coverage.
@codecov

codecov Bot commented Jul 15, 2026

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 87.32394% with 18 lines in your changes missing coverage. Please review.

Files with missing lines Patch % Lines
...oxy/management_helpers/team_metadata_validation.py 88.78% 12 Missing ⚠️
...tellm/proxy/management_endpoints/team_endpoints.py 72.22% 5 Missing ⚠️
litellm/proxy/proxy_server.py 90.00% 1 Missing ⚠️

📢 Thoughts on this report? Let us know!

if isinstance(data.metadata, dict):
TeamMemberBudgetHandler.strip_system_managed_metadata_keys(data.metadata)

await validate_team_metadata_if_configured(

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Medium: Metadata fields bypass validation

A team admin can bypass the configured metadata policy by supplying top-level fields such as tags, allowed_passthrough_routes, or secret_manager_settings. These values are added to the persisted metadata later at lines 1202–1217, after this validation has completed. The update path validates before _update_metadata_fields(), and skips the validator entirely when no explicit metadata key is supplied. Materialize all metadata-backed fields first, then validate the final dictionary that will be written.

…proxy in CI

Adds the validator matrix to the proxy_store_model_in_db_tests CircleCI
job so every scenario runs full e2e against a Postgres-backed proxy. The
proxy config registers a dispatching validator that routes each request
to one of the three implementations via a metadata key and accepts
anything that does not opt in, keeping the rest of the suite unaffected.
CI starts a stand-in cost center service on the host for the HTTP-backed
implementation, reached from the container via host.docker.internal, and
the outage path targets a closed port to prove the fail-closed 503
without stopping services.
@yuneng-berri
yuneng-berri requested a review from a team July 15, 2026 06:57
@yuneng-berri

Copy link
Copy Markdown
Contributor Author

@greptile

…it forms

The team create and edit forms asked for metadata as a raw JSON blob in a
textarea buried under Additional Settings. Both forms now render a key-value
pair editor directly under the TPM/RPM limit fields, backed by a shared
MetadataKeyValueFields component. Values round-trip losslessly: non-string
values display as JSON and parse back to their typed form on save, and
JSON-ambiguous strings are quoted so their type survives the trip. The edit
form hides UI-managed keys (logging, guardrails, model rate limits, etc.)
that dedicated controls already own and re-add on save.
@yuneng-berri

Copy link
Copy Markdown
Contributor Author

@greptile review again with the UI changes

@yuneng-berri

Copy link
Copy Markdown
Contributor Author

@greptile review again with the latest changes

@devin-ai-integration

Copy link
Copy Markdown
Contributor

QA of the team metadata validation hook, the schema, and the Admin UI editor

Ran the full happy / sad / edge matrix at be8ac082c7 against a live DB-backed proxy (Postgres, no mocks) and the PR's dashboard dev server. 33 of 35 checks passed; one real bug in the new UI component and one pre-existing gap that this feature inherits are described at the bottom

Config used

Proxy started with DATABASE_URL=postgresql://.../litellm PYTHONPATH=/path/to/validator litellm --config qa_config.yaml --port 4010, with LITELLM_LICENSE set in the env (nothing secret in the config itself)

general_settings:
  master_key: sk-1234
  custom_team_metadata_validate: qa_team_metadata_validator.validate_team_metadata
  team_metadata_validation_timeout: 3
  team_metadata_validation_error_message: "Cost center validation is unavailable right now; the team was not saved. Contact FinOps."
  team_metadata_schema:
    - key: cost_center
      label: "Cost Center"
    - key: app_name
      label: "Application Name"
    - key: chargeback_mode
      label: "Chargeback Mode"

The validator next to it accepts CC-1001 / CC-1002 and drives every branch of the contract: a missing cost center is rejected with its own message, CC-KABOOM raises, CC-SLOW hangs past the 3s timeout, CC-GARBAGE returns a malformed result, and CC-HTTPEXC raises an HTTPException(403)

import asyncio
from fastapi import HTTPException
from litellm.proxy.management_helpers.team_metadata_validation import (
    TeamMetadataValidationPayload,
    TeamMetadataValidationResult,
)

VALID_COST_CENTERS = {"CC-1001", "CC-1002"}

async def validate_team_metadata(payload: TeamMetadataValidationPayload) -> TeamMetadataValidationResult:
    cost_center = payload.metadata.get("cost_center")
    if cost_center == "CC-KABOOM":
        raise RuntimeError("simulated internal validation service outage")
    if cost_center == "CC-SLOW":
        await asyncio.sleep(30)
    if cost_center == "CC-GARBAGE":
        return {"not_a_valid": "result"}
    if cost_center == "CC-HTTPEXC":
        raise HTTPException(status_code=403, detail={"error": "validator said 403"})
    if cost_center is None:
        return TeamMetadataValidationResult(valid=False, error_message="Team metadata must include a cost_center. Contact the FinOps team.")
    if not isinstance(cost_center, str) or cost_center not in VALID_COST_CENTERS:
        return TeamMetadataValidationResult(valid=False, error_message=f"Cost center {cost_center!r} is not recognized. Contact the FinOps team.")
    return TeamMetadataValidationResult(valid=True)

def sync_validate_team_metadata(payload):
    return TeamMetadataValidationResult(valid=True)

Happy paths

$ curl -X POST localhost:4010/team/new -d '{"team_alias":"qa-valid-team","team_id":"qa-team-valid-1","metadata":{"cost_center":"CC-1001"}}'
{..."metadata":{"cost_center":"CC-1001"}...}                                                    HTTP 200

$ curl -X PATCH localhost:4010/team/qa-team-valid-1 -d '{"metadata":{"team_notes":"hello"}}'
{..."metadata":{"team_notes":"hello","cost_center":"CC-1001"}...}                               HTTP 200

$ curl -X POST localhost:4010/team/update -d '{"team_id":"qa-team-valid-1","metadata":{"cost_center":"CC-1002","team_notes":"hello"}}'
HTTP 200

$ curl -X POST localhost:4010/team/update -d '{"team_id":"qa-team-valid-1","tpm_limit":50}'     # no metadata key, validator skipped
HTTP 200

$ curl -X POST localhost:4010/team/new -d '{"team_alias":"qa-nested","team_id":"qa-team-nested-1","metadata":{"cost_center":"CC-1001","nested":{"a":[1,2,{"b":true}]},"flag":false,"num":1.5}}'
{..."metadata":{"num":1.5,"flag":false,"nested":{"a":[1,2,{"b":true}]},"cost_center":"CC-1001"}...}   HTTP 200

The PATCH payload the validator saw confirms the merged result and the requester identity, and that system-managed keys are stripped from both sides

[validator] operation=update team_id=qa-team-valid-1 metadata={'cost_center': 'CC-1001', 'team_notes': 'hello'} existing_metadata={'cost_center': 'CC-1001'} requester={'user_id': 'default_user_id', 'user_email': None, 'user_role': 'proxy_admin'}

Sad paths

$ curl -X POST localhost:4010/team/new -d '{"team_alias":"qa-bad-team","team_id":"qa-team-bad-1","metadata":{"cost_center":"CC-9999"}}'
{"error":{"message":"{'error': \"Cost center 'CC-9999' is not recognized. Contact the FinOps team.\"}",...,"code":"400"}}

$ curl -X POST localhost:4010/team/new -d '{"team_alias":"qa-no-meta-team","team_id":"qa-team-nometa-1"}'          # no metadata at all
{"error":{"message":"{'error': 'Team metadata must include a cost_center. Contact the FinOps team.'}",...,"code":"400"}}

$ curl -X POST localhost:4010/team/new -d '{"team_alias":"qa-empty-meta","team_id":"qa-team-empty-1","metadata":{}}'
HTTP 400

$ curl localhost:4010/team/info?team_id=qa-team-bad-1                                            # rejected create left no row
{"error":{"message":"{'message': 'Team not found, passed team id: qa-team-bad-1.'}",...,"code":"404"}}

$ curl -X PATCH localhost:4010/team/qa-team-valid-1 -d '{"metadata":{"cost_center":null}}'       # deleting the required key
HTTP 400

$ curl -X POST localhost:4010/team/update -d '{"team_id":"qa-team-valid-1","metadata":{"team_notes":"only-notes"}}'
HTTP 400

$ curl localhost:4010/team/info?team_id=qa-team-valid-1 | jq .team_info.metadata                 # blocked writes changed nothing
{"team_notes":"hello","cost_center":"CC-1002"}

Fail-closed and misconfiguration edge cases

$ curl -X POST localhost:4010/team/new -d '{...,"metadata":{"cost_center":"CC-KABOOM"}}'     # validator raises
{"error":{"message":"{'error': 'Cost center validation is unavailable right now; the team was not saved. Contact FinOps.'}",...,"code":"503"}}

$ curl -X POST localhost:4010/team/new -d '{...,"metadata":{"cost_center":"CC-SLOW"}}'       # hangs past team_metadata_validation_timeout: 3
HTTP 503, same generic message, returned after ~3s

$ curl -X POST localhost:4010/team/new -d '{...,"metadata":{"cost_center":"CC-GARBAGE"}}'    # malformed return value
HTTP 503

$ curl -X POST localhost:4010/team/new -d '{...,"metadata":{"cost_center":"CC-HTTPEXC"}}'    # validator raises HTTPException(403)
HTTP 503, generic message; the documented contract, the validator's own status is deliberately not surfaced
# proxy started with LITELLM_LICENSE unset
$ curl -X POST localhost:4011/team/new -d '{...,"metadata":{"cost_center":"CC-1001"}}'
{"error":{"message":"{'error': 'custom_team_metadata_validate is an Enterprise feature. You must be a LiteLLM Enterprise user...'}",...,"code":"400"}}

# custom_team_metadata_validate pointed at a non-async function
$ curl -X POST localhost:4012/team/new -d '{...,"metadata":{"cost_center":"CC-1001"}}'
{"error":{"message":"{'error': 'custom_team_metadata_validate must be an async function'}",...,"code":"500"}}

# no validator and no schema configured: behaviour is exactly as before
$ curl -X POST localhost:4013/team/new -d '{...,"metadata":{"cost_center":"CC-9999"}}'
HTTP 200
$ curl localhost:4013/team/metadata_schema -H 'Authorization: Bearer sk-1234'
{"fields":[]}

A rejected create that also carried model_aliases leaves no orphan row, so the fix for that review finding holds

$ curl -X POST localhost:4010/team/new -d '{"team_alias":"qa-alias-reject","team_id":"qa-team-alias-1","metadata":{"cost_center":"CC-9999"},"model_aliases":{"gpt":"fake-openai"}}'
HTTP 400
$ docker exec qa-pg psql -U postgres -d litellm -c 'select count(*) from "LiteLLM_ModelTable";'
 count
-------
     0

Schema endpoint and startup validation

$ curl localhost:4010/team/metadata_schema -H 'Authorization: Bearer sk-1234'
{"fields":[{"key":"cost_center","label":"Cost Center"},{"key":"app_name","label":"Application Name"},{"key":"chargeback_mode","label":"Chargeback Mode"}]}

$ curl -o /dev/null -w '%{http_code}\n' localhost:4010/team/metadata_schema                        # no key
401
$ curl -o /dev/null -w '%{http_code}\n' localhost:4010/team/metadata_schema -H 'Authorization: Bearer sk-wrong'
401

# non-admin (internal_user) key, the case the route-group fix was about
$ curl localhost:4010/team/metadata_schema -H "Authorization: Bearer $INTERNAL_USER_KEY"
{"fields":[{"key":"cost_center",...}]}                                                             HTTP 200

Malformed schemas fail proxy startup rather than serving a broken form

# unknown field name
exit code: 3
pydantic_core._pydantic_core.ValidationError: 1 validation error for tuple[TeamMetadataFieldSchema, ...]
  Extra inputs are not permitted [type=extra_forbidden, input_value=True, input_type=bool]

# duplicate keys
exit code: 3
ValueError: team_metadata_schema contains duplicate keys: cost_center

# empty key
exit code: 3
0.key
  String should have at least 1 character [type=string_too_short, input_value='', input_type=str]

Admin UI

Dashboard dev server on port 3000 (NEXT_PUBLIC_BASE_URL=http://localhost:4010 npm run dev) against the same proxy

Create Team metadata editor

The Metadata section renders as key-value rows directly under the RPM field, with no JSON textarea in Additional Settings, and is prepopulated from GET /team/metadata_schema

Fail-closed toast

The rest of the UI matrix
  • CC-9999 on create: blocked with Error creating the team: Cost center 'CC-9999' is not recognized. Contact the FinOps team., no ApiError: prefix and no {'error': ...} wrapper, and no row in /team/list
  • cost_center row removed: Error creating the team: Team metadata must include a cost_center. Contact the FinOps team.
  • CC-KABOOM: Error creating the team: Cost center validation is unavailable right now; the team was not saved. Contact FinOps.
  • a second cost_center row: inline Duplicate key, the request never fires
  • typed values round-trip: limits entered as {"a":1} is stored as a JSON object and redisplayed as {"a":1} in the edit form
  • edit form prefills stored keys and leaks no UI-managed keys into the pair editor; saving CC-9999 is blocked with Failed to update team settings: ... and the form stays open; CC-1002 saves and /team/info shows the new value with the other keys untouched

Edit form prefill
Update rejected
Duplicate key

Bug: blank metadata values are persisted as empty strings

Creating a team with only cost_center filled in, leaving the schema-seeded app_name and chargeback_mode rows untouched, saves them as empty strings

$ curl -sS 'http://localhost:4010/team/info?team_id=0e75256e-...' -H 'Authorization: Bearer sk-1234' | jq .team_info.metadata
{"limits": {"a": 1}, "app_name": "", "cost_center": "CC-1001", "chargeback_mode": ""}

The same happens on edit: clearing a stored value writes "limits": "" instead of removing the key. metadataPairsToObject in MetadataKeyValueFields.tsx keeps every pair that has a non-empty key, so dropping pairs whose value is empty or whitespace-only would fix both. This matters beyond tidiness because the validator now receives "" for schema keys the admin never filled in, so a validator written as if metadata.get("cost_center") is None: reject passes on an empty string. The PR description claims blank fields are omitted rather than saved as empty strings, so this looks like a regression against intent rather than a design choice

Pre-existing gap the feature inherits: metadata-backed top-level fields bypass the validator

An update that carries no metadata key but does carry a metadata-backed field such as tags skips validation and then overwrites the stored metadata wholesale, dropping the validated cost center

$ curl -X POST localhost:4010/team/update -d '{"team_id":"qa-team-valid-1","tags":["sneaky"]}'
HTTP 200
$ curl localhost:4010/team/info?team_id=qa-team-valid-1 | jq .team_info.metadata
{"tags":["sneaky"]}

The wholesale overwrite is pre-existing; I reproduced the same result on litellm_internal_staging with no validator configured, so it is not caused by this PR. It does mean the enforcement promise has a hole: a team admin can clear a required metadata key with a tags-only update. This is the same thing the veria-ai comment flagged, and it is worth either fixing here (validate the final materialized metadata rather than only updated_kv["metadata"]) or writing down as a known limitation. On create the same fields are merged in after validation, though there the merge is additive so the required key survives

@yuneng-berri

Copy link
Copy Markdown
Contributor Author

@greptile A validator that raises HTTPException(403) does get converted to the generic 503 - this is intentional

review again

…itellm_/metadata-prepopulation-design-b107cb

# Conflicts:
#	litellm/types/proxy/management_endpoints/team_endpoints.py
#	tests/test_litellm/proxy/management_endpoints/test_team_endpoints.py
@yuneng-berri

Copy link
Copy Markdown
Contributor Author

@veria-ai review

@yuneng-berri
yuneng-berri merged commit cd87fee into litellm_internal_staging Aug 4, 2026
80 of 81 checks passed
@yuneng-berri
yuneng-berri deleted the litellm_/bold-mclaren-8d89b3 branch August 4, 2026 01:37
pull Bot pushed a commit to chizee/litellm that referenced this pull request Aug 4, 2026
The management route-coverage guard fires because /team/metadata_schema landed
in BerriAI#33353 without a behavior-suite scenario, so this adds one covering the nine
seeded actors plus the unauthenticated 401

The prometheus budget-metric assertions read the log call's first positional
arg, which BerriAI#35703 turned into an unrendered "%s" format string when it moved
logging to lazy args. They now render the message from the call args, which
also pins the arg order and the exception text that the old substring check
never reached

GitHub Models was fully retired on 2026-07-30, so test_completion_github_api
can no longer pass: the endpoint the github provider targets returns 404 and
models.github.ai answers 410 "github_models_retirement_brownout". The dead live
test is removed rather than skipped
doonga pushed a commit to greyrock-labs/home-ops that referenced this pull request Aug 17, 2026
…7.0) (#336)

This PR contains the following updates:

| Package | Update | Change |
|---|---|---|
| [ghcr.io/berriai/litellm](https://images.chainguard.dev/directory/image/wolfi-base/overview) ([source](https://github.com/BerriAI/litellm)) | minor | `v1.96.2` → `v1.97.0` |

---

### Release Notes

<details>
<summary>BerriAI/litellm (ghcr.io/berriai/litellm)</summary>

### [`v1.97.0`](https://github.com/BerriAI/litellm/releases/tag/v1.97.0)

[Compare Source](https://github.com/BerriAI/litellm/compare/v1.97.0...v1.97.0)

##### Verify Docker Image Signature

All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0).

**Verify using the pinned commit hash (recommended):**

A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \
  ghcr.io/berriai/litellm:v1.97.0
```

**Verify using the release tag (convenience):**

Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/v1.97.0/cosign.pub \
  ghcr.io/berriai/litellm:v1.97.0
```

Expected output:

```
The following checks were performed on each of these signatures:
  - The cosign claims were validated
  - The signatures were verified against the specified public key
```

***

##### What's Changed

- feat(proxy): resolve Cursor thinking/fast model-name suffixes on /cursor/chat/completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35554](https://github.com/BerriAI/litellm/pull/35554)
- fix(team-callbacks): actually stop logging when disable\_logging is called by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35520](https://github.com/BerriAI/litellm/pull/35520)
- refactor(lint): drop redundant !s f-string conversion flags and fix displaced import-group comments by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35546](https://github.com/BerriAI/litellm/pull/35546)
- fix(proxy): backfill null user\_email on existing users during JWT auth by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34588](https://github.com/BerriAI/litellm/pull/34588)
- feat(playground): add non-streaming response toggle by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35560](https://github.com/BerriAI/litellm/pull/35560)
- feat(teams): apply default organization to new teams from default team settings by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35540](https://github.com/BerriAI/litellm/pull/35540)
- fix(ui): block Playground page for viewer roles on direct URL access by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35676](https://github.com/BerriAI/litellm/pull/35676)
- fix(caching): close evicted LLM clients so their connections are reclaimed by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35492](https://github.com/BerriAI/litellm/pull/35492)
- chore(deps): update brace-expansion, postcss, and gitpython to current patch releases by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35692](https://github.com/BerriAI/litellm/pull/35692)
- refactor(ui): rename the create MCP server component to PascalCase by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35686](https://github.com/BerriAI/litellm/pull/35686)
- fix(openai): drop undefined Union from owns\_wrapped\_http\_client annotation by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35706](https://github.com/BerriAI/litellm/pull/35706)
- fix(openai): drop the undefined Union from owns\_wrapped\_http\_client by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35704](https://github.com/BerriAI/litellm/pull/35704)
- chore(ui): note Google's Agent Platform rename in vector store setup by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;28076](https://github.com/BerriAI/litellm/pull/28076)
- fix(proxy): apply key/team router\_settings.model\_group\_alias by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35486](https://github.com/BerriAI/litellm/pull/35486)
- feat(complexity\_router): default session affinity off and expose it in the UI by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35714](https://github.com/BerriAI/litellm/pull/35714)
- fix(datadog): read team callback dd\_\* params from kwargs instead of blocked dynamic params ([#&#8203;35115](https://github.com/BerriAI/litellm/issues/35115) port) by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35687](https://github.com/BerriAI/litellm/pull/35687)
- refactor(ui): extract the MCP create form's logic and field groups by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35694](https://github.com/BerriAI/litellm/pull/35694)
- test(ui): tier the MCP create tests into unit and integration by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35697](https://github.com/BerriAI/litellm/pull/35697)
- fix(proxy): redact credential headers from request logging copies by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35678](https://github.com/BerriAI/litellm/pull/35678)
- feat(guardrails/rubrik): prompt moderation, response-text blocking, streaming buffer, failure logging by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35722](https://github.com/BerriAI/litellm/pull/35722)
- fix(ui): render Responses API request and response in the logs drawer by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35718](https://github.com/BerriAI/litellm/pull/35718)
- fix(ui): hide guardrail review buttons from non-admin users by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;27535](https://github.com/BerriAI/litellm/pull/27535)
- feat(team): custom metadata validation hook for team create and update by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;33353](https://github.com/BerriAI/litellm/pull/33353)
- ci(circleci): install a pinned Rust toolchain on the Linux jobs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35519](https://github.com/BerriAI/litellm/pull/35519)
- fix(bedrock): stop forwarding no-op toolSpec.strict to Converse by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35688](https://github.com/BerriAI/litellm/pull/35688)
- fix(ui): reject an auto-router keyword rule left empty instead of dropping it by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35705](https://github.com/BerriAI/litellm/pull/35705)
- fix(guardrails/rubrik): attribute blocked requests to the caller that made them by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35734](https://github.com/BerriAI/litellm/pull/35734)
- fix(responses): forward client headers to the provider on /v1/responses by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34531](https://github.com/BerriAI/litellm/pull/34531)
- feat(spend): add net auto-router savings to the cost-optimization dashboard by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35521](https://github.com/BerriAI/litellm/pull/35521)
- chore(typing): clear basedpyright Any errors in budget reset, access groups, and cache settings by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35719](https://github.com/BerriAI/litellm/pull/35719)
- fix(spend): read what a request cost from the record instead of pricing it again by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35736](https://github.com/BerriAI/litellm/pull/35736)
- perf: install hiredis so redis-py parses replies with its C parser by [@&#8203;Classic298](https://github.com/Classic298) in [#&#8203;35709](https://github.com/BerriAI/litellm/pull/35709)
- feat(ui): show auto-router savings on the cost-optimization dashboard by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35522](https://github.com/BerriAI/litellm/pull/35522)
- perf: build log messages lazily so filtered-out log records cost nothing by [@&#8203;Classic298](https://github.com/Classic298) in [#&#8203;35703](https://github.com/BerriAI/litellm/pull/35703)
- fix(proxy): retry model cost map fetch with Retry-After-aware backoff and keep current map on reload failure by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35739](https://github.com/BerriAI/litellm/pull/35739)
- feat(otel): stamp service tier attributes on inference spans by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35679](https://github.com/BerriAI/litellm/pull/35679)
- fix(proxy): log the model cost map reload failure lazily by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35750](https://github.com/BerriAI/litellm/pull/35750)
- fix(groq): translate web\_search\_options to the browser\_search tool by [@&#8203;hMED22](https://github.com/hMED22) in [#&#8203;34971](https://github.com/BerriAI/litellm/pull/34971)
- feat(ui): add admin-configurable user banner by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35729](https://github.com/BerriAI/litellm/pull/35729)
- fix(e2e): make spend-counter redis connection env-driven for non-cluster deployments by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35732](https://github.com/BerriAI/litellm/pull/35732)
- fix(proxy): make /cursor/chat/completions work with Cursor agent mode by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;34029](https://github.com/BerriAI/litellm/pull/34029)
- fix(proxy): propagate user\_email and bind api\_key on JWT auth attribution paths by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34331](https://github.com/BerriAI/litellm/pull/34331)
- chore(build): move the Admin UI toolchain to Node 24 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35801](https://github.com/BerriAI/litellm/pull/35801)
- test(e2e): vendor API strategy coverage across endpoints by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;34649](https://github.com/BerriAI/litellm/pull/34649)
- chore(deps): upgrade cryptography to 50.0.0 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35803](https://github.com/BerriAI/litellm/pull/35803)
- test(e2e): cover legacy text /completions endpoint by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;34431](https://github.com/BerriAI/litellm/pull/34431)
- feat(gemini): add gemini-robotics-er-2-preview and gemini-robotics-er-1.6-preview by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35555](https://github.com/BerriAI/litellm/pull/35555)
- test(e2e): move load/perf testing out of the main suite and drop the vllm passthrough test by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35820](https://github.com/BerriAI/litellm/pull/35820)
- feat(lint): enforce Final on locals and freeze function parameters (LIT010, LIT011) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35807](https://github.com/BerriAI/litellm/pull/35807)
- chore: bump litellm-proxy-extras 0.4.81 -> 0.4.82, litellm 1.96.0 -> 1.97.0 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35810](https://github.com/BerriAI/litellm/pull/35810)
- fix(bedrock): drop conflicting tool\_choice.type when toolConfig.toolChoice is set by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35738](https://github.com/BerriAI/litellm/pull/35738)
- docs(CLAUDE.md): prefer commas over semicolons when replacing em dashes by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35825](https://github.com/BerriAI/litellm/pull/35825)
- chore(lint): zero out basedpyright headroom for purely local rules by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35828](https://github.com/BerriAI/litellm/pull/35828)
- test(e2e): retry provider-transient statuses at the transport with bounded backoff by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35824](https://github.com/BerriAI/litellm/pull/35824)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35836](https://github.com/BerriAI/litellm/pull/35836)
- refactor(ui): route MCP session tokens through the shared storage helper by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35835](https://github.com/BerriAI/litellm/pull/35835)
- docs(helm): replace the classic chart's 128Mi resource example with the documented 4Gi sizing by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35830](https://github.com/BerriAI/litellm/pull/35830)
- fix(proxy): persist periodic reload schedule state so status survives restarts and fires without store\_model\_in\_db by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35165](https://github.com/BerriAI/litellm/pull/35165)
- fix(router): eagerly fetch Vertex AI deferred stream to surface HTTP errors in \_acompletion fallback path by [@&#8203;deepanshululla](https://github.com/deepanshululla) in [#&#8203;34627](https://github.com/BerriAI/litellm/pull/34627)
- fix(azure\_storage): honor AZURE\_STORAGE\_ENDPOINT\_SUFFIX for sovereign clouds by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35806](https://github.com/BerriAI/litellm/pull/35806)
- fix(proxy): apply key\_alias/key\_hash filters to all /key/list visibility branches by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35840](https://github.com/BerriAI/litellm/pull/35840)
- fix(proxy): enforce per-model budgets against resolved cursor model variants by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35834](https://github.com/BerriAI/litellm/pull/35834)
- feat(ui): reorder Add Auto Router into name + template, with a collapsible detailed config by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35746](https://github.com/BerriAI/litellm/pull/35746)
- test: repair three failing suites on litellm\_internal\_staging by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35845](https://github.com/BerriAI/litellm/pull/35845)
- fix(guardrails): scan model output on the /openai/v1/responses alias by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35818](https://github.com/BerriAI/litellm/pull/35818)
- ci: pin Node on the Playwright UI lanes so npm ci meets the engines floor by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35848](https://github.com/BerriAI/litellm/pull/35848)
- fix(pricing): apply OpenAI's gpt-5.6 terra/luna cut to Azure cost map by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;35481](https://github.com/BerriAI/litellm/pull/35481)
- feat(spend): add caller-scoped key/user/team/organization spend report endpoints by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35725](https://github.com/BerriAI/litellm/pull/35725)
- revert: "fix(caching): close evicted LLM clients so their connections are reclaimed ([#&#8203;35492](https://github.com/BerriAI/litellm/issues/35492))" by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35856](https://github.com/BerriAI/litellm/pull/35856)
- refactor(repositories): add prisma protocol seams and a spend-reset unit of work by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35748](https://github.com/BerriAI/litellm/pull/35748)
- perf(streaming): assemble streamed tool-call arguments in linear time by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35826](https://github.com/BerriAI/litellm/pull/35826)
- fix(s3\_v2): sign S3 object URLs with S3SigV4Auth so encoded paths verify by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35726](https://github.com/BerriAI/litellm/pull/35726)
- test(e2e): self-seed the ui suite's password-login users in global setup by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35863](https://github.com/BerriAI/litellm/pull/35863)
- fix(claude-code): create-only skill registration with a PUT update route (LIT-4110) by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;31752](https://github.com/BerriAI/litellm/pull/31752)
- fix(proxy):  fix zguard  httpcode when block input by [@&#8203;jwang-gif](https://github.com/jwang-gif) in [#&#8203;31948](https://github.com/BerriAI/litellm/pull/31948)
- fix(lint): pick the merge-aware base so in-progress merges are not blamed for base drift by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35868](https://github.com/BerriAI/litellm/pull/35868)
- chore: bump litellm-proxy-extras 0.4.82 -> 0.4.83 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35877](https://github.com/BerriAI/litellm/pull/35877)
- feat(ui): add Test Routing to the auto router create form by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35859](https://github.com/BerriAI/litellm/pull/35859)
- fix(ui): derive auto-router preset tests from the bundled preset JSON by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35882](https://github.com/BerriAI/litellm/pull/35882)
- revert: "test(e2e): vendor API strategy coverage across endpoints" ([#&#8203;34649](https://github.com/BerriAI/litellm/issues/34649)) by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35881](https://github.com/BerriAI/litellm/pull/35881)
- chore(deps): bump grpc and golang.org/x modules in the terraform provider by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35844](https://github.com/BerriAI/litellm/pull/35844)
- test(e2e): skip view-backed global spend probes pending LIT-5211 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35875](https://github.com/BerriAI/litellm/pull/35875)
- fix(lint): move the basedpyright heap flag into the type check gate by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35869](https://github.com/BerriAI/litellm/pull/35869)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35876](https://github.com/BerriAI/litellm/pull/35876)
- feat(ui): add role capability gating, migrate Tool Policies route by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35812](https://github.com/BerriAI/litellm/pull/35812)
- refactor(ui): inject the fetch client's base url instead of reading it at import by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35802](https://github.com/BerriAI/litellm/pull/35802)
- chore: remove unused .flake8 config and flake8 dev dependency by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35888](https://github.com/BerriAI/litellm/pull/35888)
- chore: stop advising pre-commit and bootstrap by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35884](https://github.com/BerriAI/litellm/pull/35884)
- fix(auth): name enable\_jwt\_auth when a JWT-shaped key is rejected by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35831](https://github.com/BerriAI/litellm/pull/35831)
- feat(auto-router): make reminder marker pair configurable by [@&#8203;akapur99](https://github.com/akapur99) in [#&#8203;35874](https://github.com/BerriAI/litellm/pull/35874)
- fix(UI): update anthropic model presets by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35896](https://github.com/BerriAI/litellm/pull/35896)
- fix(bootstrap): switch to the dashboard node floor via nvm or fnm by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35895](https://github.com/BerriAI/litellm/pull/35895)
- perf(pre-commit): run python, dashboard, and gen-api checks concurrently by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35903](https://github.com/BerriAI/litellm/pull/35903)
- feat(spend): derive a default auto-router savings baseline from the hardest tier by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35907](https://github.com/BerriAI/litellm/pull/35907)
- fix(http\_handler): self-heal handler clients closed after cache eviction by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35862](https://github.com/BerriAI/litellm/pull/35862)
- fix(cost\_tracking): keep OpenAI prompt cache token details through usage reassembly by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34812](https://github.com/BerriAI/litellm/pull/34812)
- fix(cost): bill gpt-5.6 prompt cache reads at the cache read rate by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34957](https://github.com/BerriAI/litellm/pull/34957)
- fix(batches): account for Responses API usage by [@&#8203;rimysore](https://github.com/rimysore) in [#&#8203;35367](https://github.com/BerriAI/litellm/pull/35367)
- ci: retry Codecov uploads and stop failing jobs on OIDC token flakes by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35251](https://github.com/BerriAI/litellm/pull/35251)
- feat(complexity\_router): let operators rename the four complexity tiers by [@&#8203;akapur99](https://github.com/akapur99) in [#&#8203;35893](https://github.com/BerriAI/litellm/pull/35893)
- chore(lint): zero stale ruff and LIT headroom and strip inert type: ignore comments by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35928](https://github.com/BerriAI/litellm/pull/35928)
- chore(lint): zero out seven more purely local basedpyright rules by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35927](https://github.com/BerriAI/litellm/pull/35927)
- chore(ui): zero stale headroom on local dashboard eslint budgets by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35929](https://github.com/BerriAI/litellm/pull/35929)
- fix(managed-files): skip rows without file objects by [@&#8203;rimysore](https://github.com/rimysore) in [#&#8203;35365](https://github.com/BerriAI/litellm/pull/35365)
- fix(router): redact fallback tracebacks at the call site and cover the sync deferred stream by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35843](https://github.com/BerriAI/litellm/pull/35843)
- fix(migrations): recover from an interrupted Prisma toolchain install by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35832](https://github.com/BerriAI/litellm/pull/35832)
- fix(lint): bring basedpyright rule counts back under their budget limits by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35962](https://github.com/BerriAI/litellm/pull/35962)
- chore(ui): don't zero out stale headroom except no-console by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35964](https://github.com/BerriAI/litellm/pull/35964)
- fix(proxy): give proxy\_admin\_viewer read parity with proxy\_admin by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35851](https://github.com/BerriAI/litellm/pull/35851)
- refactor(ui): address UI lint budget issues by refactoring UI by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35960](https://github.com/BerriAI/litellm/pull/35960)
- fix(ci): make the env-key doc gate see get\_secret\_bool reads by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35833](https://github.com/BerriAI/litellm/pull/35833)
- fix(caching): re-land evicted LLM client closing ([#&#8203;35492](https://github.com/BerriAI/litellm/issues/35492)) atop self-healing handlers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35870](https://github.com/BerriAI/litellm/pull/35870)
- fix(proxy): keep the connected DB client when a startup health check fails by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35837](https://github.com/BerriAI/litellm/pull/35837)
- chore(lint): remove litellm/types from the ruff lint exclusion by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35926](https://github.com/BerriAI/litellm/pull/35926)
- feat(sgr): make the gateway middleware the source of truth for successful requests by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35717](https://github.com/BerriAI/litellm/pull/35717)
- feat(auto-router): let operators replace the LLM classifier's system prompt by [@&#8203;akapur99](https://github.com/akapur99) in [#&#8203;35855](https://github.com/BerriAI/litellm/pull/35855)
- fix(docker): bake the pip image's prisma engines at a world-readable path by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35976](https://github.com/BerriAI/litellm/pull/35976)
- fix(auth): return 403 from the OAuth2 enterprise gate by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35838](https://github.com/BerriAI/litellm/pull/35838)
- fix(router): keep custom model\_info across a price data reload by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35491](https://github.com/BerriAI/litellm/pull/35491)
- fix(proxy): resolve pass-through credentials live from router deployments by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35916](https://github.com/BerriAI/litellm/pull/35916)
- fix(ci): fetch only head and merge-base in lint jobs instead of every branch by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35982](https://github.com/BerriAI/litellm/pull/35982)
- fix(autorouter): match CJK keyword\_tier\_rules that regex word boundaries miss by [@&#8203;akapur99](https://github.com/akapur99) in [#&#8203;35984](https://github.com/BerriAI/litellm/pull/35984)
- feat(spend): rebuild the auto-router benchmarks backend as a per-session rollup by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35910](https://github.com/BerriAI/litellm/pull/35910)
- refactor(ui): replace hand-rolled query-param routing with nuqs by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35871](https://github.com/BerriAI/litellm/pull/35871)
- fix(docker): bake the componentized prisma engines at /opt/prisma so any uid can start by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35989](https://github.com/BerriAI/litellm/pull/35989)
- fix(migrations): keep the toolchain heal from raising on an unreadable nodeenv cache by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35986](https://github.com/BerriAI/litellm/pull/35986)
- fix(bedrock): sign Bedrock managed-file S3 requests with S3SigV4Auth by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35983](https://github.com/BerriAI/litellm/pull/35983)
- chore(typing): replace Any seams with real types across responses, proxy, and provider adapters by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35809](https://github.com/BerriAI/litellm/pull/35809)
- fix(ai21): resolve the documented AI21\_API\_KEY instead of a misspelled name by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35985](https://github.com/BerriAI/litellm/pull/35985)
- fix(docker): fail the image build when the generated prisma engine paths drift off /opt/prisma by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35979](https://github.com/BerriAI/litellm/pull/35979)
- fix(jina\_ai): resolve the documented JINA\_API\_KEY as a fallback by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35992](https://github.com/BerriAI/litellm/pull/35992)
- fix(proxy): only treat a recoverable database outage as grounds to serve without one by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35864](https://github.com/BerriAI/litellm/pull/35864)
- fix(ci): make every remaining CI checkout shallow by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35997](https://github.com/BerriAI/litellm/pull/35997)
- fix(auto-router): stop the embedding model's context window from failing long requests by [@&#8203;akapur99](https://github.com/akapur99) in [#&#8203;35956](https://github.com/BerriAI/litellm/pull/35956)
- fix(ci): make the env-key doc gate see bare get\_secret and get\_secret\_str reads by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35996](https://github.com/BerriAI/litellm/pull/35996)
- fix(logging): extend secret redaction to records litellm does not emit directly by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35977](https://github.com/BerriAI/litellm/pull/35977)
- test(utils): pin the register\_model replay test to the recorded half by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35994](https://github.com/BerriAI/litellm/pull/35994)
- fix(ci): run every helm test suite, not just the first one per file by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35993](https://github.com/BerriAI/litellm/pull/35993)
- ci: fail the build when a test file or Dockerfile is invoked by no job by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35991](https://github.com/BerriAI/litellm/pull/35991)
- fix(langfuse): stop a collected httpx handler from closing a shared client by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35981](https://github.com/BerriAI/litellm/pull/35981)
- fix(bedrock): grant bedrock:CountTokens in OIDC session policy by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;33145](https://github.com/BerriAI/litellm/pull/33145)
- feat(pre-commit): save full lint output to a per-worktree log file by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36004](https://github.com/BerriAI/litellm/pull/36004)
- feat(ui): match auto-router preset models against deployments' underlying model IDs by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35972](https://github.com/BerriAI/litellm/pull/35972)
- fix(core\_helpers): map generic 'error' finish\_reason to 'stop' by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;33972](https://github.com/BerriAI/litellm/pull/33972)
- fix(proxy)!: apply request-parameter checks consistently across body, path and form inputs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36011](https://github.com/BerriAI/litellm/pull/36011)
- fix: rebuild models\_by\_provider in add\_known\_models so cost map reloads reach wildcard expansion by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36010](https://github.com/BerriAI/litellm/pull/36010)
- feat(complexity\_router): report LLM classifier cost per request via routing\_decision and x-litellm-classifier-cost header by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36015](https://github.com/BerriAI/litellm/pull/36015)
- fix(model-prices): correct replicate model key typo by [@&#8203;AkashNaickar](https://github.com/AkashNaickar) in [#&#8203;34800](https://github.com/BerriAI/litellm/pull/34800)
- fix(proxy): register managed batch output files on terminal retrieve by [@&#8203;Souravrajvi0](https://github.com/Souravrajvi0) in [#&#8203;34092](https://github.com/BerriAI/litellm/pull/34092)
- perf(pre-commit): fetch basedpyright base counts from CI artifacts by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35970](https://github.com/BerriAI/litellm/pull/35970)
- fix(ui): sync projects list page index to ?page= so back and reload keep the page by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36003](https://github.com/BerriAI/litellm/pull/36003)
- fix(ui): link project page keys to their virtual key detail by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36002](https://github.com/BerriAI/litellm/pull/36002)
- refactor(ui): drop unreferenced locals from dashboard route components by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35819](https://github.com/BerriAI/litellm/pull/35819)
- fix(ui): opening a project now pushes ?project= so back and deep links work by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36001](https://github.com/BerriAI/litellm/pull/36001)
- refactor(ui): drop unreferenced locals from shared dashboard components by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35821](https://github.com/BerriAI/litellm/pull/35821)
- refactor(ui): drop unreferenced locals from tests and narrow destructures by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36025](https://github.com/BerriAI/litellm/pull/36025)
- fix(guardrails): allow litellm\_content\_filter to run on post\_mcp\_call by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35980](https://github.com/BerriAI/litellm/pull/35980)
- fix(guardrails): scan /v1/messages tool traffic by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35999](https://github.com/BerriAI/litellm/pull/35999)
- refactor(ui): drop dead locals and unused React state across the dashboard by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36026](https://github.com/BerriAI/litellm/pull/36026)
- feat(ui): add the auto-router usage tab to cost optimization by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35995](https://github.com/BerriAI/litellm/pull/35995)
- fix(managed\_files): derive unified output file ids deterministically so concurrent registrations converge by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36019](https://github.com/BerriAI/litellm/pull/36019)
- fix(proxy): send keepalive pings on anthropic messages SSE streams during upstream silence by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36024](https://github.com/BerriAI/litellm/pull/36024)
- fix(managed\_files): return unified ids from unscoped file listing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36031](https://github.com/BerriAI/litellm/pull/36031)
- fix(arize\_phoenix): lowercase OTLP/gRPC auth metadata key by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34883](https://github.com/BerriAI/litellm/pull/34883)
- fix(auto-router): accept every reminder marker pair a harness emits by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36029](https://github.com/BerriAI/litellm/pull/36029)
- fix(pricing): sync flex/priority tier keys to dated OpenAI snapshot variants by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35923](https://github.com/BerriAI/litellm/pull/35923)
- fix(cost): bill reasoning tokens at the service tier output rate by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35925](https://github.com/BerriAI/litellm/pull/35925)
- fix(proxy): include today's UTC bucket when a daily activity range ends at the caller's current day by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36051](https://github.com/BerriAI/litellm/pull/36051)
- fix: expired-miss share over all measured turns + cost-optimization tab labels by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36037](https://github.com/BerriAI/litellm/pull/36037)
- fix(router): include Bedrock batch/S3 fields and model in deployment credentials by [@&#8203;mpcusack-altos](https://github.com/mpcusack-altos) in [#&#8203;24548](https://github.com/BerriAI/litellm/pull/24548)
- fix(batch): track cost for managed batches with no attributable key/u… by [@&#8203;elinacse](https://github.com/elinacse) in [#&#8203;35468](https://github.com/BerriAI/litellm/pull/35468)
- feat(guardrails): add scan\_only\_tool\_results to scope unified guardrails to tool results by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36014](https://github.com/BerriAI/litellm/pull/36014)
- fix(cost): stop token-pricing the placeholder input on file content calls by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35140](https://github.com/BerriAI/litellm/pull/35140)
- fix(proxy): fetch background responses through the router in CheckResponsesCost by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35137](https://github.com/BerriAI/litellm/pull/35137)
- fix(proxy): yaml store\_prompts\_in\_spend\_logs should take precedence over DB cached value by [@&#8203;Praveena-617](https://github.com/Praveena-617) in [#&#8203;35769](https://github.com/BerriAI/litellm/pull/35769)
- fix(lint): measure the basedpyright budget gate in a gate-owned venv by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36050](https://github.com/BerriAI/litellm/pull/36050)
- docs: cap all GitHub comments at 15-25 words, curb semicolon splices by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36059](https://github.com/BerriAI/litellm/pull/36059)
- chore(lint): name MappingProxyType in the mutable-collection fix messages by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36072](https://github.com/BerriAI/litellm/pull/36072)
- test: roll back runtime model registrations between tests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36039](https://github.com/BerriAI/litellm/pull/36039)
- refactor(types): cut 653 implicit and explicit Any diagnostics across 11 modules by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36054](https://github.com/BerriAI/litellm/pull/36054)
- fix(proxy): stop resolving the UI session sentinel team on /search\_tools/list by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36061](https://github.com/BerriAI/litellm/pull/36061)
- fix(batches): persist managed file ids for cancelled/failed/expired batches by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36048](https://github.com/BerriAI/litellm/pull/36048)
- fix(batches): register managed output files on batch cancel by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36034](https://github.com/BerriAI/litellm/pull/36034)
- fix(proxy): allow non-admins to reach /user/daily/activity/aggregated by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36062](https://github.com/BerriAI/litellm/pull/36062)
- fix(anthropic): coerce explicit additionalProperties to false in output\_format schema by [@&#8203;dkindlund](https://github.com/dkindlund) in [#&#8203;35811](https://github.com/BerriAI/litellm/pull/35811)
- fix(batches): prevent managed file fallbacks by [@&#8203;rimysore](https://github.com/rimysore) in [#&#8203;35371](https://github.com/BerriAI/litellm/pull/35371)
- chore: ignore the mechanical lint and typing sweeps in git blame by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36076](https://github.com/BerriAI/litellm/pull/36076)
- fix(proxy): warn at startup when max\_budget is set but no database is connected by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36041](https://github.com/BerriAI/litellm/pull/36041)
- fix(proxy): promote caller metadata trace fields into litellm\_metadata by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35866](https://github.com/BerriAI/litellm/pull/35866)
- feat(terraform): sync provider 0.3.0 from the mirror and cut 0.4.0 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36098](https://github.com/BerriAI/litellm/pull/36098)
- fix(guardrails): honor configured timeout in Zscaler AI Guard by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36110](https://github.com/BerriAI/litellm/pull/36110)
- fix(logging): fall back to litellm\_metadata when metadata is empty by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36105](https://github.com/BerriAI/litellm/pull/36105)
- fix(proxy): re-assert the authenticated identity on passthrough requests by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36121](https://github.com/BerriAI/litellm/pull/36121)
- chore: bump litellm-enterprise 0.1.53 -> 0.1.54, litellm-proxy-extras 0.4.83 -> 0.4.84 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36139](https://github.com/BerriAI/litellm/pull/36139)
- fix(ui): match auto-router preset models against wildcard-expanded model groups by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36111](https://github.com/BerriAI/litellm/pull/36111)
- test(router): assert the auto-router max\_input\_chars kwarg by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36109](https://github.com/BerriAI/litellm/pull/36109)
- fix(ui): allow clearing a key's budget reset from the Edit Key form by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36140](https://github.com/BerriAI/litellm/pull/36140)
- fix(managed\_files): skip unparseable rows when listing managed files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36021](https://github.com/BerriAI/litellm/pull/36021)
- fix(a2a): stop writing per-caller headers onto the shared cached httpx client by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35978](https://github.com/BerriAI/litellm/pull/35978)
- build(deps): bump h2 to 4.4.1 and js-yaml to 4.3.1 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36147](https://github.com/BerriAI/litellm/pull/36147)
- chore: promote staging to main by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36057](https://github.com/BerriAI/litellm/pull/36057)
- fix(azure\_sentinel): respect AZURE\_AUTHORITY\_HOST and derive the Azure Monitor audience per cloud by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36137](https://github.com/BerriAI/litellm/pull/36137)
- fix(bedrock): pass SSE-KMS key through to the batch input-file S3 upload by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35148](https://github.com/BerriAI/litellm/pull/35148)
- fix(anthropic adapter): stop indexing choices\[0] on choiceless streaming chunks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35314](https://github.com/BerriAI/litellm/pull/35314)
- fix(bedrock): normalize /v1/completions and /v1/responses batch records by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35675](https://github.com/BerriAI/litellm/pull/35675)
- fix(proxy): return the real status code when a credential update is rejected by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36166](https://github.com/BerriAI/litellm/pull/36166)
- fix(proxy): improve Headroom /v1/compress HTTP 404 diagnostics by [@&#8203;aayush598](https://github.com/aayush598) in [#&#8203;35952](https://github.com/BerriAI/litellm/pull/35952)
- fix(proxy): invalidate cached project object on project update and delete by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36028](https://github.com/BerriAI/litellm/pull/36028)
- feat(proxy): add apply\_user\_budget\_to\_team\_keys opt-in by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36102](https://github.com/BerriAI/litellm/pull/36102)
- fix(proxy): stop alerting on health probes that lose the planned engine-restart race by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;36141](https://github.com/BerriAI/litellm/pull/36141)
- test(docker): gate the componentized gateway and backend images on an arbitrary-uid offline boot by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;36136](https://github.com/BerriAI/litellm/pull/36136)
- fix(http): stop pooled clients persisting cookies on the aiohttp jar too by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;36149](https://github.com/BerriAI/litellm/pull/36149)
- fix(router): bound fallback-walk work and error-log volume by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;36148](https://github.com/BerriAI/litellm/pull/36148)
- ci: wire credential\_endpoints tests into the proxy endpoints job by [@&#8203;cursor](https://github.com/cursor)\[bot] in [#&#8203;36187](https://github.com/BerriAI/litellm/pull/36187)
- docs(keys): document /key/info fields and clarify budget\_reset\_at is the next reset by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36127](https://github.com/BerriAI/litellm/pull/36127)
- fix(azure\_sentinel): add AZURE\_SENTINEL\_AUTHORITY\_HOST as a Sentinel scoped override by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36165](https://github.com/BerriAI/litellm/pull/36165)
- docs(pr-template): add a User Flow section with authoring instructions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36162](https://github.com/BerriAI/litellm/pull/36162)
- fix(proxy): derive config agent ids from agent\_name so grants survive secret rotation by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36020](https://github.com/BerriAI/litellm/pull/36020)
- chore(ui): regenerate schema.d.ts for the /key/info docstring update by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36210](https://github.com/BerriAI/litellm/pull/36210)
- build(deps): bump gitpython to 3.1.58 to clear osv-scan on staging by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36212](https://github.com/BerriAI/litellm/pull/36212)
- fix(proxy): deny agent access when key and team grants resolve to nothing by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36221](https://github.com/BerriAI/litellm/pull/36221)
- build(deps): defer the second pypdf advisory until the 6.15.0 bump by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36218](https://github.com/BerriAI/litellm/pull/36218)
- fix(a2a): align agent list annotation and test with the tuple return type by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36217](https://github.com/BerriAI/litellm/pull/36217)
- ci: always run the UI API types sync check so it can be required by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36213](https://github.com/BerriAI/litellm/pull/36213)
- build(deps): bump nanoid to 3.3.17 in the dashboard lockfile by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36227](https://github.com/BerriAI/litellm/pull/36227)
- feat(ui): show user email or alias in usage data export by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36232](https://github.com/BerriAI/litellm/pull/36232)
- feat(auto-router): track turns per complexity tier (LIT-5302) by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36209](https://github.com/BerriAI/litellm/pull/36209)
- fix(websearch): restore snippet text in native web\_search\_tool\_result blocks (LIT-5315) by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36228](https://github.com/BerriAI/litellm/pull/36228)
- fix(proxy): resolve entity access groups in the model listing endpoints by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36230](https://github.com/BerriAI/litellm/pull/36230)
- fix(ui): let access groups be a team's only model source, with hover provenance by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;36234](https://github.com/BerriAI/litellm/pull/36234)
- fix(managed\_files): return unified output file ids from GET /batches by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36049](https://github.com/BerriAI/litellm/pull/36049)
- test(proxy): compare empty agent list to the tuple get\_agent\_list returns by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36225](https://github.com/BerriAI/litellm/pull/36225)
- fix(otel): name the RPC system and upstream on MCP tool-call spans by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35857](https://github.com/BerriAI/litellm/pull/35857)
- fix(guardrails): chunk oversized Bedrock ApplyGuardrail requests instead of failing by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;36119](https://github.com/BerriAI/litellm/pull/36119)
- test(e2e): settle control-plane writes across every replica, not just one by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36247](https://github.com/BerriAI/litellm/pull/36247)
- fix(responses): forward allowed\_openai\_params through the chat completions bridge by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35885](https://github.com/BerriAI/litellm/pull/35885)
- test(proxy): assert the copy \_add\_team\_member\_budget\_table returns by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36244](https://github.com/BerriAI/litellm/pull/36244)
- chore(ui): regenerate dashboard api types for tier\_turns by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36243](https://github.com/BerriAI/litellm/pull/36243)
- refactor(types): declare mirrored pricing fields on ModelInfo by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36215](https://github.com/BerriAI/litellm/pull/36215)
- fix(lint): make strict-gate noqas survive base ruff and flag stale ones by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36257](https://github.com/BerriAI/litellm/pull/36257)
- fix(vertex\_ai): surface real error/status on vertex batch create instead of IndexError 500 by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35141](https://github.com/BerriAI/litellm/pull/35141)
- ci: give the remaining pull\_request workflows a concurrency group by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36252](https://github.com/BerriAI/litellm/pull/36252)
- refactor(lint): graduate zero-violation strict rules and guard the budget ratchet by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36161](https://github.com/BerriAI/litellm/pull/36161)
- fix(proxy): enforce require\_managed\_files on every route that accepts a raw provider id by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35551](https://github.com/BerriAI/litellm/pull/35551)
- chore(typing): clear 1.4k basedpyright Any errors across 21 hotspot files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36282](https://github.com/BerriAI/litellm/pull/36282)
- test: roll back live router replay membership between tests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36278](https://github.com/BerriAI/litellm/pull/36278)
- chore(ci): sync main into internal staging by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36288](https://github.com/BerriAI/litellm/pull/36288)
- build(lint): rename make pre-commit to make check with a working-tree fallback by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36277](https://github.com/BerriAI/litellm/pull/36277)
- fix(ui): show team BYOK models in team fallback settings by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36241](https://github.com/BerriAI/litellm/pull/36241)
- fix(otel): mark v2 server spans as failed for pre-call errors by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34546](https://github.com/BerriAI/litellm/pull/34546)
- fix(websearch\_interception): bill intercepted searches to the calling key by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35708](https://github.com/BerriAI/litellm/pull/35708)
- chore: remove pre-commit rule by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36295](https://github.com/BerriAI/litellm/pull/36295)
- docs: clarify guideline priority ordering in CLAUDE.md by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36296](https://github.com/BerriAI/litellm/pull/36296)
- feat(router): independent, default-on deployment affinity for the auto-router by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;36146](https://github.com/BerriAI/litellm/pull/36146)
- test: repair stale CircleCI contracts by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36293](https://github.com/BerriAI/litellm/pull/36293)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36286](https://github.com/BerriAI/litellm/pull/36286)
- chore: rebuild Admin UI bundle for the 2026-08-08 release by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36297](https://github.com/BerriAI/litellm/pull/36297)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;36304](https://github.com/BerriAI/litellm/pull/36304)

##### New Contributors

- [@&#8203;rimysore](https://github.com/rimysore) made their first contribution in [#&#8203;35367](https://github.com/BerriAI/litellm/pull/35367)
- [@&#8203;AkashNaickar](https://github.com/AkashNaickar) made their first contribution in [#&#8203;34800](https://github.com/BerriAI/litellm/pull/34800)
- [@&#8203;Souravrajvi0](https://github.com/Souravrajvi0) made their first contribution in [#&#8203;34092](https://github.com/BerriAI/litellm/pull/34092)
- [@&#8203;elinacse](https://github.com/elinacse) made their first contribution in [#&#8203;35468](https://github.com/BerriAI/litellm/pull/35468)
- [@&#8203;aayush598](https://github.com/aayush598) made their first contribution in [#&#8203;35952](https://github.com/BerriAI/litellm/pull/35952)
- [@&#8203;cursor](https://github.com/cursor)\[bot] made their first contribution in [#&#8203;36187](https://github.com/BerriAI/litellm/pull/36187)

**Full Changelog**: <https://github.com/BerriAI/litellm/compare/v1.96.0...v1.97.0>

### [`v1.97.0`](https://github.com/BerriAI/litellm/releases/tag/v1.97.0)

[Compare Source](https://github.com/BerriAI/litellm/compare/v1.96.2...v1.97.0)

##### Verify Docker Image Signature

All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0).

**Verify using the pinned commit hash (recommended):**

A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \
  ghcr.io/berriai/litellm:v1.97.0
```

**Verify using the release tag (convenience):**

Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/v1.97.0/cosign.pub \
  ghcr.io/berriai/litellm:v1.97.0
```

Expected output:

```
The following checks were performed on each of these signatures:
  - The cosign claims were validated
  - The signatures were verified against the specified public key
```

***

##### What's Changed

- feat(proxy): resolve Cursor thinking/fast model-name suffixes on /cursor/chat/completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35554](https://github.com/BerriAI/litellm/pull/35554)
- fix(team-callbacks): actually stop logging when disable\_logging is called by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35520](https://github.com/BerriAI/litellm/pull/35520)
- refactor(lint): drop redundant !s f-string conversion flags and fix displaced import-group comments by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35546](https://github.com/BerriAI/litellm/pull/35546)
- fix(proxy): backfill null user\_email on existing users during JWT auth by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34588](https://github.com/BerriAI/litellm/pull/34588)
- feat(playground): add non-streaming response toggle by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35560](https://github.com/BerriAI/litellm/pull/35560)
- feat(teams): apply default organization to new teams from default team settings by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35540](https://github.com/BerriAI/litellm/pull/35540)
- fix(ui): block Playground page for viewer roles on direct URL access by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35676](https://github.com/BerriAI/litellm/pull/35676)
- fix(caching): close evicted LLM clients so their connections are reclaimed by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35492](https://github.com/BerriAI/litellm/pull/35492)
- chore(deps): update brace-expansion, postcss, and gitpython to current patch releases by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35692](https://github.com/BerriAI/litellm/pull/35692)
- refactor(ui): rename the create MCP server component to PascalCase by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35686](https://github.com/BerriAI/litellm/pull/35686)
- fix(openai): drop undefined Union from owns\_wrapped\_http\_client annotation by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35706](https://github.com/BerriAI/litellm/pull/35706)
- fix(openai): drop the undefined Union from owns\_wrapped\_http\_client by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35704](https://github.com/BerriAI/litellm/pull/35704)
- chore(ui): note Google's Agent Platform rename in vector store setup by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;28076](https://github.com/BerriAI/litellm/pull/28076)
- fix(proxy): apply key/team router\_settings.model\_group\_alias by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35486](https://github.com/BerriAI/litellm/pull/35486)
- feat(complexity\_router): default session affinity off and expose it in the UI by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35714](https://github.com/BerriAI/litellm/pull/35714)
- fix(datadog): read team callback dd\_\* params from kwargs instead of blocked dynamic params ([#&#8203;35115](https://github.com/BerriAI/litellm/issues/35115) port) by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35687](https://github.com/BerriAI/litellm/pull/35687)
- refactor(ui): extract the MCP create form's logic and field groups by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35694](https://github.com/BerriAI/litellm/pull/35694)
- test(ui): tier the MCP create tests into unit and integration by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35697](https://github.com/BerriAI/litellm/pull/35697)
- fix(proxy): redact credential headers from request logging copies by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35678](https://github.com/BerriAI/litellm/pull/35678)
- feat(guardrails/rubrik): prompt moderation, response-text blocking, streaming buffer, failure logging by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35722](https://github.com/BerriAI/litellm/pull/35722)
- fix(ui): render Responses API request and response in the logs drawer by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35718](https://github.com/BerriAI/litellm/pull/35718)
- fix(ui): hide guardrail review buttons from non-admin users by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;27535](https://github.com/BerriAI/litellm/pull/27535)
- feat(team): custom metadata validation hook for team create and update by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;33353](https://github.com/BerriAI/litellm/pull/33353)
- ci(circleci): install a pinned Rust toolchain on the Linux jobs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35519](https://github.com/BerriAI/litellm/pull/35519)
- fix(bedrock): stop forwarding no-op toolSpec.strict to Converse by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35688](https://github.com/BerriAI/litellm/pull/35688)
- fix(ui): reject an auto-router keyword rule left empty instead of dropping it by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35705](https://github.com/BerriAI/litellm/pull/35705)
- fix(guardrails/rubrik): attribute blocked requests to the caller that made them by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;35734](https://github.com/BerriAI/litellm/pull/35734)
- fix(responses): forward client headers to the provider on /v1/responses by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34531](https://github.com/BerriAI/litellm/pull/34531)
- feat(spend): add net auto-router savings to the cost-optimization dashboard by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35521](https://github.com/BerriAI/litellm/pull/35521)
- chore(typing): clear basedpyright Any errors in budget reset, access groups, and cache settings by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;35719](https://github.com/BerriAI/litellm/pull/35719)
- fix(spend): read what a request cost from the record instead of pricing it again by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35736](https://github.com/BerriAI/litellm/pull/35736)
- perf: install hiredis so redis-py parses replies with its C parser by [@&#8203;Classic298](https://github.com/Classic298) in [#&#8203;35709](https://github.com/BerriAI/litellm/pull/35709)
- feat(ui): show auto-router savings on the cost-optimization dashboard by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35522](https://github.com/BerriAI/litellm/pull/35522)
- perf: build log messages lazily so filtered-out log records cost nothing by [@&#8203;Classic298](https://github.com/Classic298) in [#&#8203;35703](https://github.com/BerriAI/litellm/pull/35703)
- fix(proxy): retry model cost map fetch with Retry-After-aware backoff and keep current map on reload failure by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;35739](https://github.com/BerriAI/litellm/pull/35739)
- feat(otel): stamp service tier attributes on inference spans by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35679](https://github.com/BerriAI/litellm/pull/35679)
- fix(proxy): log the model cost map reload failure lazily by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;35750](https://github.com/BerriAI/litellm/pull/35750)
- fix(groq): translate web\_search\_options to the browser\_search tool by [@&#8203;hMED22](https://github.com/hMED22) in [#&#8203;34971](https://github.com/BerriAI/litellm/pull/34971)
- feat(ui): add admin-configurable user banner by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35729](https://github.com/BerriAI/litellm/pull/35729)
- fix(e2e): make spend-counter redis connection env-driven for non-cluster deployments by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;35732](https://github.com/BerriAI/litellm/pull/35732)
- fix(proxy): make /cursor/chat/completions work with Cursor agent mode by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;34029](https://github.com/BerriAI/litellm/pull/34029)
- fix(proxy): propagate user\_email and bind api\_key on JWT auth attribution paths by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;34331](https://github.com/BerriAI/litellm/pull/34331)
- chore(build): move the Admin UI toolchain to Node 24 by [@&#8203;yuneng-berri](https://g…
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants