Skip to content

feat(key management): show budget window usage on /key/info - #37044

Merged
ryan-crabbe-berri merged 15 commits into
BerriAI:litellm_internal_stagingfrom
Thijmen:key-budget-window-usage
Sep 1, 2026
Merged

ryan-crabbe-berri merged 15 commits into
BerriAI:litellm_internal_stagingfrom
Thijmen:key-budget-window-usage

Conversation

@Thijmen

@Thijmen Thijmen commented Aug 15, 2026 •

Copy link
Copy Markdown
Contributor

TLDR

Problem this solves:

  • A key's budget windows show the limits but never how much is used
  • The only usage signal today is the 429 once a window runs out

How it solves it:

  • /key/info and /v2/key/info add a budget_limits_usage field keyed by budget_duration with current_spend per window, a sibling of budget_limits in the same spirit as model_max_budget_usage
  • budget_limits comes back exactly as stored, so nothing that round-trips it into /key/update ever sees a computed field
  • The number comes from the same counter and LiteLLM_BudgetWindowSpend row enforcement reads, not a spend-log aggregate

User Flow

Before: an admin who put an hourly and a daily budget on a key cannot see how close either window is to its cap

  1. They send POST https://litellm-domain/key/generate with "budget_limits": [{"budget_duration": "1h", "max_budget": 2.0}, {"budget_duration": "1d", "max_budget": 20.0}] and get back an sk-... key
  2. Their app sends POST https://litellm-domain/v1/chat/completions with that key and gets a 200 with x-litellm-response-cost: 0.000295
  3. They send GET https://litellm-domain/key/info?key=sk-... and each budget_limits entry shows only reset_at, max_budget, and budget_duration
  4. The first time they learn a window is used up is a 429 ExceededBudget: Key over 1h budget on a real request

After: the same admin sees per-window usage on the key

  1. They send POST https://litellm-domain/key/generate with the same budget_limits and get back an sk-... key
  2. Their app sends POST https://litellm-domain/v1/chat/completions with that key and gets a 200 with x-litellm-response-cost: 0.000295
  3. They send GET https://litellm-domain/key/info?key=sk-... and the response now carries "budget_limits_usage": {"1h": {"current_spend": 0.0003}, "1d": {"current_spend": 0.0003}} next to the unchanged budget_limits
  4. POST https://litellm-domain/v2/key/info with {"keys": ["sk-..."]} returns the same budget_limits_usage on every key

Relevant issues

Linear ticket

Pre-Submission checklist

Please complete all items before asking a LiteLLM maintainer to review your PR

  • I have added meaningful tests
  • The handful of test files covering my change pass locally, e.g. uv run pytest tests/test_litellm/<your_test_file>.py -v. Leave the suites (make test-unit-*, make test-unit) to CI: it finishes in ~15 minutes where a laptop takes an hour or more
  • My PR passes all required CI/CD checks (e.g., lint, schema.d.ts sync check, etc.)
  • My PR's scope is as isolated as possible; it only solves 1 specific problem
  • I have received a Greptile Confidence Score of at least 4/5 before requesting a maintainer review (Greptile reviews automatically once the PR is opened; only comment @greptileai to re-request a review after pushing changes)

Delays in PR merge?

If you're seeing a delay in your PR being merged, ping the LiteLLM Team on Slack (#pr-review).

Screenshots / Proof of Fix

Shared setup: litellm/proxy/dev_config.yaml (master key sk-1234, gpt-5.5 backed by a real OpenAI key), one Postgres for both runs. The After proxy runs on :4373, the Before proxy (staging tip 3fadcd7155) on :4374. The same key is used on both sides, created and driven on the After proxy first

curl -s -X POST http://localhost:4373/key/generate -H "Authorization: Bearer sk-1234" -H 'Content-Type: application/json' \
  -d '{"key_alias":"pr37044-window-qa","models":["gpt-5.5"],"budget_limits":[{"budget_duration":"1h","max_budget":2.0},{"budget_duration":"1d","max_budget":20.0}]}'
# => {"key":"sk-ziYZm4U4ErrktAbIvenbvQ", ...}

curl -s -D - -o /dev/null http://localhost:4373/v1/chat/completions -H "Authorization: Bearer sk-ziYZm4U4ErrktAbIvenbvQ" -H 'Content-Type: application/json' \
  -d '{"model":"gpt-5.5","messages":[{"role":"user","content":"Say hi in three words"}]}' | grep -i "^HTTP/\|x-litellm-response-cost:"
# HTTP/1.1 200 OK
# x-litellm-response-cost: 0.000295

Before (3fadcd7)

  1. curl -s "http://localhost:4374/key/info?key=sk-ziYZm4U4ErrktAbIvenbvQ" -H "Authorization: Bearer sk-1234" | jq '.info | {budget_limits, budget_limits_usage}'

    {
      "budget_limits": [
        { "reset_at": "2026-09-01T05:00:00+00:00", "max_budget": 2.0, "budget_duration": "1h" },
        { "reset_at": "2026-09-02T00:00:00+00:00", "max_budget": 20.0, "budget_duration": "1d" }
      ],
      "budget_limits_usage": null
    }
  2. curl -s -X POST http://localhost:4374/v2/key/info -H "Authorization: Bearer sk-1234" -H 'Content-Type: application/json' -d '{"keys":["sk-ziYZm4U4ErrktAbIvenbvQ"]}' | jq '.info[0] | {budget_limits, budget_limits_usage}'

    {
      "budget_limits": [
        { "reset_at": "2026-09-01T05:00:00+00:00", "max_budget": 2.0, "budget_duration": "1h" },
        { "reset_at": "2026-09-02T00:00:00+00:00", "max_budget": 20.0, "budget_duration": "1d" }
      ],
      "budget_limits_usage": null
    }

After (760b864)

  1. curl -s "http://localhost:4373/key/info?key=sk-ziYZm4U4ErrktAbIvenbvQ" -H "Authorization: Bearer sk-1234" | jq '.info | {budget_limits, budget_limits_usage}' (before any traffic)

    {
      "budget_limits": [
        { "reset_at": "2026-09-01T05:00:00+00:00", "max_budget": 2.0, "budget_duration": "1h" },
        { "reset_at": "2026-09-02T00:00:00+00:00", "max_budget": 20.0, "budget_duration": "1d" }
      ],
      "budget_limits_usage": {
        "1h": { "current_spend": 0.0 },
        "1d": { "current_spend": 0.0 }
      }
    }
  2. Same /key/info call as step 1, about 15 seconds after the chat completion above

    {
      "budget_limits": [
        { "reset_at": "2026-09-01T05:00:00+00:00", "max_budget": 2.0, "budget_duration": "1h" },
        { "reset_at": "2026-09-02T00:00:00+00:00", "max_budget": 20.0, "budget_duration": "1d" }
      ],
      "budget_limits_usage": {
        "1h": { "current_spend": 0.0003 },
        "1d": { "current_spend": 0.0003 }
      }
    }
  3. curl -s -X POST http://localhost:4373/v2/key/info -H "Authorization: Bearer sk-1234" -H 'Content-Type: application/json' -d '{"keys":["sk-ziYZm4U4ErrktAbIvenbvQ"]}' | jq '.info[0] | {budget_limits, budget_limits_usage}'

    {
      "budget_limits": [
        { "reset_at": "2026-09-01T05:00:00+00:00", "max_budget": 2.0, "budget_duration": "1h" },
        { "reset_at": "2026-09-02T00:00:00+00:00", "max_budget": 20.0, "budget_duration": "1d" }
      ],
      "budget_limits_usage": {
        "1h": { "current_spend": 0.0003 },
        "1d": { "current_spend": 0.0003 }
      }
    }
  4. The window spend rows the read is backed by, once the writer flushed (psql "$DATABASE_URL" -c 'select entity_type, window_duration, window_start, spend from "LiteLLM_BudgetWindowSpend" where entity_id = <sha256 of the key> order by window_duration')

     entity_type | window_duration |    window_start     |  spend
    -------------+-----------------+---------------------+----------
     key         | 1d              | 2026-09-01 00:00:00 | 0.000295
     key         | 1h              | 2026-09-01 04:00:00 | 0.000295
    
  5. Proof the read is the table, not a spend-log aggregate: poison the 1h row, wait 6s for the floor cache, then re-read (psql "$DATABASE_URL" -c 'update "LiteLLM_BudgetWindowSpend" set spend = 999 where window_duration = '"'"'1h'"'"' and entity_id = <sha256 of the key>', then the same /key/info call as step 1)

    {
      "budget_limits": [
        { "reset_at": "2026-09-01T05:00:00+00:00", "max_budget": 2.0, "budget_duration": "1h" },
        { "reset_at": "2026-09-02T00:00:00+00:00", "max_budget": 20.0, "budget_duration": "1d" }
      ],
      "budget_limits_usage": {
        "1h": { "current_spend": 999.0 },
        "1d": { "current_spend": 0.0003 }
      }
    }

Type

🆕 New Feature

Caveats (if any)

Medium

  • /v2/key/info now does one counter read per window per key, so large batches of windowed keys cost more than before

Low

  • The Admin UI does not display budget_limits_usage yet; only the API response carries it

Final Attestation

  • The tests check the right things, including the edge cases, and regressions in the respective real-world customer use-cases are not possible after this PR

@greptile-apps

greptile-apps Bot commented Aug 15, 2026 •

Copy link
Copy Markdown
Contributor

Greptile Summary

The PR adds current budget-window spend information to the single-key and batched key-information APIs while preserving the stored budget_limits shape.

  • Builds usage entries from the budget enforcement counters.
  • Adds tests for single-key, batched, empty, JSON-backed, and model-backed budget-window configurations.
  • Updates the generated dashboard API schema documentation.

Confidence Score: 5/5

The PR appears safe to merge.

No blocking failure remains.

Important Files Changed

Filename Overview
litellm/proxy/management_endpoints/key_management_endpoints.py Adds immutable budget-window usage mappings and attaches them to both key-information endpoint responses.
tests/test_litellm/proxy/management_endpoints/test_key_management_endpoints.py Adds focused unit coverage for budget-window usage retrieval and supported stored input shapes.
ui/litellm-dashboard/src/lib/http/schema.d.ts Documents the additional key-information response fields in the generated API declarations.

Reviews (6): Last reviewed commit: "refactor(key): trim budget_limits_usage ..." | Re-trigger Greptile

Comment thread litellm/proxy/management_endpoints/key_management_endpoints.py Outdated
Comment thread litellm/proxy/management_endpoints/key_management_endpoints.py Outdated
@codecov

codecov Bot commented Aug 15, 2026 •

Copy link
Copy Markdown

Codecov Report

❌ Patch coverage is 90.00000% with 3 lines in your changes missing coverage. Please review.

Files with missing lines Patch % Lines
...y/management_endpoints/key_management_endpoints.py 90.00% 3 Missing ⚠️

📢 Thoughts on this report? Let us know!

@codspeed

codspeed Bot commented Aug 15, 2026 •

Copy link
Copy Markdown
Contributor

Merging this PR will not alter performance

✅ 31 untouched benchmarks


Comparing Thijmen:key-budget-window-usage (760b864) with litellm_internal_staging (d132040)

Open in CodSpeed

Comment thread litellm/proxy/management_endpoints/key_management_endpoints.py Outdated
@veria-ai

veria-ai Bot commented Aug 15, 2026 •

Copy link
Copy Markdown
Contributor

PR overview

All previously flagged issues have been addressed. No open security concerns remain on this pull request.

Security review

No open security issues remain on this pull request.

Fixed/addressed: 1 · PR risk: 0/10

Replace _attach_budget_limits_usage, which rewrote the caller's key_info dict, with _budget_limits_with_usage returning a new list. Callers assign the result once. Keeps the response shape and spend-counter read path identical while following the repo's no-mutation rule.
@Thijmen

Thijmen commented Aug 17, 2026

Copy link
Copy Markdown
Contributor Author

@greptile-apps

1 similar comment
@Thijmen

Thijmen commented Aug 17, 2026

Copy link
Copy Markdown
Contributor Author

@greptile-apps

…ey-budget-window-usage

# Conflicts:
#	tests/test_litellm/proxy/management_endpoints/test_key_management_endpoints.py
The two `or []` fallbacks only feed len(), so tuples do the job without
a mutable literal. The detail dict mirrors the sibling HTTPException
payload above and is never mutated, so it carries a mutable-ok reason.
caching-local, proxy-extras and enterprise-package gave pytest 20m but
capped the job at 55m; with a 35m setup ceiling plus 5m of runner
overhead the job deadline could preempt pytest itself.
@Thijmen
Thijmen requested a review from a team August 24, 2026 07:09
reruns: 2
timeout-minutes: 20
job-timeout-minutes: 55
job-timeout-minutes: 60

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

I dont think this change needs to be here

@ryan-crabbe-berri

ryan-crabbe-berri commented Sep 1, 2026 •

Copy link
Copy Markdown
Contributor

Thanks for this. I'm taking it over to rebase onto staging, where I refactored the way we do spend tracking for budget windows. Prev version did a pretty expensive query on the spend logs table

detail={"message": "Malformed request. No keys passed in."},
)
requested_key_count: Final = len(data.keys or ()) + len(data.key_aliases or ())
if requested_key_count > MAX_KEY_INFO_KEYS_PER_REQUEST:

@ryan-crabbe-berri ryan-crabbe-berri Sep 1, 2026 •

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

this is a pretty big breaking behavior change, will revert because i refactored the budget windows spend tracking

…able

Pass window_duration to get_current_spend so /key/info re-checks a stale-low counter against the LiteLLM_BudgetWindowSpend row instead of aggregating LiteLLM_SpendLogs, and reuse _budget_limit_windows for the stored-column coercion. Drop the /v2/key/info batch cap (a new 422 for callers that work today) and the unrelated CI timeout bump and soft_budget test
@ryan-crabbe-berri ryan-crabbe-berri changed the title feat: add show budget window usage feat(key management): show budget window usage on /key/info Sep 1, 2026
@ryan-crabbe-berri

Copy link
Copy Markdown
Contributor

@greptile re review

…of inlining current_spend

budget_limits now comes back exactly as stored on /key/info and /v2/key/info.
The per-window usage moves to a sibling budget_limits_usage field keyed by
budget_duration (current_spend, budget_limit, reset_at), mirroring
model_max_budget_usage, so the stored shape that /key/update accepts never
carries a computed field.
@ryan-crabbe-berri

Copy link
Copy Markdown
Contributor

@greptile re review

max_budget and reset_at already live on the matching budget_limits entry, so
repeating them (as budget_limit and reset_at) only invited confusion about which
copy is authoritative.
@ryan-crabbe-berri

Copy link
Copy Markdown
Contributor

@greptile re review

@ryan-crabbe-berri ryan-crabbe-berri left a comment

Copy link
Copy Markdown
Contributor

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Thanks!

@ryan-crabbe-berri
ryan-crabbe-berri merged commit fa720be into BerriAI:litellm_internal_staging Sep 1, 2026
79 checks passed
doonga pushed a commit to greyrock-labs/home-ops that referenced this pull request Sep 15, 2026
…101.0) (#201)

This PR contains the following updates:

| Package | Update | Change |
|---|---|---|
| [ghcr.io/berriai/litellm](https://images.chainguard.dev/directory/image/wolfi-base/overview) ([source](https://github.com/BerriAI/litellm)) | minor | `v1.100.1` → `v1.101.0` |

---

### Release Notes

<details>
<summary>BerriAI/litellm (ghcr.io/berriai/litellm)</summary>

### [`v1.101.0`](https://github.com/BerriAI/litellm/releases/tag/v1.101.0)

[Compare Source](https://github.com/BerriAI/litellm/compare/v1.100.1...v1.101.0)

#### Verify Docker Image Signature

All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0).

**Verify using the pinned commit hash (recommended):**

A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \
  ghcr.io/berriai/litellm:v1.101.0
```

**Verify using the release tag (convenience):**

Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/v1.101.0/cosign.pub \
  ghcr.io/berriai/litellm:v1.101.0
```

Expected output:

```
The following checks were performed on each of these signatures:
  - The cosign claims were validated
  - The signatures were verified against the specified public key
```

***

#### What's Changed

- fix(proxy): emit timing headers and overhead for /v1/messages and /v1/responses by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38840](https://github.com/BerriAI/litellm/pull/38840)
- fix(tests): derive the no-cache-read-rate savings baseline from the model map by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38863](https://github.com/BerriAI/litellm/pull/38863)
- chore(typing): clear Any seams across 47 files, ratchet basedpyright ceilings -3,302 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;37778](https://github.com/BerriAI/litellm/pull/37778)
- chore(typing): clear 1.2k basedpyright Any errors across 16 hotspot files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36722](https://github.com/BerriAI/litellm/pull/36722)
- feat(bedrock): honor streaming buffer/sampling config for unbuffered post\_call scans by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38722](https://github.com/BerriAI/litellm/pull/38722)
- feat(cli): set ENABLE\_TOOL\_SEARCH=true for lite claude by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38942](https://github.com/BerriAI/litellm/pull/38942)
- fix(proxy): deliver budget alerts on webhook-only alerting and accept ALERTING\_WEBHOOK\_URL by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38441](https://github.com/BerriAI/litellm/pull/38441)
- docs(claude.md): require tests to check behavior, not code structure by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;38772](https://github.com/BerriAI/litellm/pull/38772)
- chore(newrelic): cover static default\_team\_settings per-team routing by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38857](https://github.com/BerriAI/litellm/pull/38857)
- fix: update stale source URLs and deprecation dates in model cost map by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38801](https://github.com/BerriAI/litellm/pull/38801)
- feat(ci): close duplicate issues after a 3-day grace period by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;38381](https://github.com/BerriAI/litellm/pull/38381)
- docs(proxy): clarify spend semantics on /v2/user/info and /user/daily/activity by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38883](https://github.com/BerriAI/litellm/pull/38883)
- fix(guardrails): configure Prompt Security file timeout policy by [@&#8203;davida-ps](https://github.com/davida-ps) in [#&#8203;38083](https://github.com/BerriAI/litellm/pull/38083)
- fix(bedrock): stop duplicating Converse config blocks inside inferenceConfig by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38993](https://github.com/BerriAI/litellm/pull/38993)
- fix(guardrails): exclude images from HiddenLayer v1 scans by [@&#8203;Ashton-Sidhu](https://github.com/Ashton-Sidhu) in [#&#8203;29210](https://github.com/BerriAI/litellm/pull/29210)
- feat(spend\_tracking): persist router metadata in spend logs for internal router models by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39001](https://github.com/BerriAI/litellm/pull/39001)
- fix(vertex\_ai): graft default vertex path when api\_base has a version-only path by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38986](https://github.com/BerriAI/litellm/pull/38986)
- fix(proxy): allow unblocking customers via /customer/update by [@&#8203;cat0825](https://github.com/cat0825) in [#&#8203;34696](https://github.com/BerriAI/litellm/pull/34696)
- feat(openai): support workload identity federation (OIDC token exchange) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38995](https://github.com/BerriAI/litellm/pull/38995)
- fix(otel): emit cache token counts on OTel v2 LLM spans by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38716](https://github.com/BerriAI/litellm/pull/38716)
- feat(proxy): add /v1/responses/input\_tokens token counting endpoint by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38997](https://github.com/BerriAI/litellm/pull/38997)
- fix(docker): bump wolfi-base for glibc 2.44 and pin apk python to 3.13 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38917](https://github.com/BerriAI/litellm/pull/38917)
- fix(docker): bump wolfi-base for glibc 2.44 and pin apk python to 3.13 in migrations image by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38973](https://github.com/BerriAI/litellm/pull/38973)
- feat(friendli): add zai-org/GLM-5.3-Flash model pricing by [@&#8203;Lee-Si-Yoon](https://github.com/Lee-Si-Yoon) in [#&#8203;38880](https://github.com/BerriAI/litellm/pull/38880)
- chore(techdebt): clear fresh debt from the 2026-08-29 and 2026-08-30 windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38884](https://github.com/BerriAI/litellm/pull/38884)
- fix(bedrock): surface Nova Sonic user transcripts, speech events, and usage in realtime API by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38597](https://github.com/BerriAI/litellm/pull/38597)
- fix(guardrails): carry Anthropic url image sources through to guardrails by [@&#8203;samtsai15](https://github.com/samtsai15) in [#&#8203;38940](https://github.com/BerriAI/litellm/pull/38940)
- feat(friendli): add zai-org/GLM-5.3 model pricing by [@&#8203;Lee-Si-Yoon](https://github.com/Lee-Si-Yoon) in [#&#8203;38881](https://github.com/BerriAI/litellm/pull/38881)
- fix(router): apply model renames to the in-memory deployment list by [@&#8203;yatishgoel](https://github.com/yatishgoel) in [#&#8203;38479](https://github.com/BerriAI/litellm/pull/38479)
- test(e2e): cover SCIM token creation and SCIM API auth in the Admin UI suite by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39027](https://github.com/BerriAI/litellm/pull/39027)
- feat(gigachat): add native API passthrough routes with spend logging by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38913](https://github.com/BerriAI/litellm/pull/38913)
- feat(gigachat): add passthrough gigachat route by [@&#8203;KnyazSh](https://github.com/KnyazSh) in [#&#8203;25886](https://github.com/BerriAI/litellm/pull/25886)
- feat(complexity-router): add classification\_mode to skip classifier on continuation turns by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38861](https://github.com/BerriAI/litellm/pull/38861)
- fix(proxy): preserve model table columns on master key rotation by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38878](https://github.com/BerriAI/litellm/pull/38878)
- fix(speech): stop forwarding response\_format as a chat param for Gemini TTS by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38819](https://github.com/BerriAI/litellm/pull/38819)
- fix(proxy): return 200 from /model/block and /model/unblock instead of 500 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38873](https://github.com/BerriAI/litellm/pull/38873)
- feat(complexity\_router): escalate oversized prompts to a tier that fits before dispatch by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38844](https://github.com/BerriAI/litellm/pull/38844)
- feat(shadow\_eval): target teams and users so JWT-auth traffic can be evaluated by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39015](https://github.com/BerriAI/litellm/pull/39015)
- fix(anthropic\_messages): drain upstream in a detached pump so client … by [@&#8203;nuernber](https://github.com/nuernber) in [#&#8203;36008](https://github.com/BerriAI/litellm/pull/36008)
- refactor(proxy): bound the budget window seed by time instead of request ids by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;38851](https://github.com/BerriAI/litellm/pull/38851)
- fix(proxy): ship psycopg so partitioned SpendLogs detection actually runs by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;38994](https://github.com/BerriAI/litellm/pull/38994)
- test(e2e): assert user-observable behavior instead of DOM structure by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39016](https://github.com/BerriAI/litellm/pull/39016)
- build(rust): configure native extension profiles by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39020](https://github.com/BerriAI/litellm/pull/39020)
- fix(ui): keep litellm\_credential\_name from LiteLLM Params JSON when no credential is selected by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39005](https://github.com/BerriAI/litellm/pull/39005)
- Revert "fix(ui): keep litellm\_credential\_name from LiteLLM Params JSON when no credential is selected" by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39046](https://github.com/BerriAI/litellm/pull/39046)
- fix(auth): quiet malformed virtual key rejections to stdout by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38838](https://github.com/BerriAI/litellm/pull/38838)
- fix(proxy): wire team-level logging callbacks into passthrough endpoints by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38979](https://github.com/BerriAI/litellm/pull/38979)
- feat(complexity\_router): opt-in modality-based capability routing for image requests by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39032](https://github.com/BerriAI/litellm/pull/39032)
- fix(ui): let the auto-router scoring tier list follow the theme by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39040](https://github.com/BerriAI/litellm/pull/39040)
- feat(ui): auto-router controls for context-window escalation by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39054](https://github.com/BerriAI/litellm/pull/39054)
- fix(redis): coerce env var string types and fix param discovery through decorator wrappers by [@&#8203;koladefaj](https://github.com/koladefaj) in [#&#8203;30644](https://github.com/BerriAI/litellm/pull/30644)
- feat(key management): show budget window usage on /key/info by [@&#8203;Thijmen](https://github.com/Thijmen) in [#&#8203;37044](https://github.com/BerriAI/litellm/pull/37044)
- fix(websearch): reject invalid explicit search tool selections by [@&#8203;georgeatparallel](https://github.com/georgeatparallel) in [#&#8203;38113](https://github.com/BerriAI/litellm/pull/38113)
- feat(shadow\_eval): compare several auto-routers on one job's sampled traffic by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39028](https://github.com/BerriAI/litellm/pull/39028)
- fix(speech): honor pcm/wav response\_format for Gemini TTS and reject unsupported containers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38868](https://github.com/BerriAI/litellm/pull/38868)
- fix(proxy): match /v1/audio/speech content-type to the returned audio format by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38798](https://github.com/BerriAI/litellm/pull/38798)
- test(e2e): drop the two mgmt registry cells no shared-proxy test can cover by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39055](https://github.com/BerriAI/litellm/pull/39055)
- feat(ui): one classification frequency picker for complexity auto-routers by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39042](https://github.com/BerriAI/litellm/pull/39042)
- test(e2e/ui): automate 8 manual QA checklist flows by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39025](https://github.com/BerriAI/litellm/pull/39025)
- fix(key\_management): allow non-admin key\_type preset transitions on /key/update by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39051](https://github.com/BerriAI/litellm/pull/39051)
- chore(typing): clear 1.1k basedpyright Any errors across 53 backend files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38796](https://github.com/BerriAI/litellm/pull/38796)
- fix(openai): forward reasoning\_effort for unknown model aliases instead of failing closed by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39065](https://github.com/BerriAI/litellm/pull/39065)
- test(e2e-ui): poll credential availability before Test Connect to deflake multi-instance runs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39073](https://github.com/BerriAI/litellm/pull/39073)
- feat(ui): modality routing toggle on the auto-router create and edit forms by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39059](https://github.com/BerriAI/litellm/pull/39059)
- fix(proxy): include litellm\_model\_table in GET /v2/team/list by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39045](https://github.com/BerriAI/litellm/pull/39045)
- fix(bedrock): mask signed request headers in guardrail debug log by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39044](https://github.com/BerriAI/litellm/pull/39044)
- fix(bedrock): forward aws\_external\_id in files and batches credential loading by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39066](https://github.com/BerriAI/litellm/pull/39066)
- fix(mcp): persist alias MCP grants verbatim instead of rewriting to local server ids by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39119](https://github.com/BerriAI/litellm/pull/39119)
- fix(responses): json-encode object tool call arguments in the chat completions bridge by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35417](https://github.com/BerriAI/litellm/pull/35417)
- fix(cost): bill OCR annotation pages via annotation\_cost\_per\_page by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38985](https://github.com/BerriAI/litellm/pull/38985)
- fix(policy\_engine): restore request guardrails list after pipeline allow by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39038](https://github.com/BerriAI/litellm/pull/39038)
- fix(embeddings): omit encoding\_format when the client omits it on OpenAI-compatible calls by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38774](https://github.com/BerriAI/litellm/pull/38774)
- test: deflake MCP registry state, savings cost map, and MCP identity env reload tests by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38891](https://github.com/BerriAI/litellm/pull/38891)
- feat(helm): add Argo CD PreSync hook and rollout strategy knobs to the componentized chart by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39112](https://github.com/BerriAI/litellm/pull/39112)
- fix(registry): veo 3.1 pricing tiers + roll up open registry PRs (glm-5.2, Qwen3.8-Flash, gemma-4-31b, scribe\_v2, fireworks/databricks deepseek v4) + deprecation dates by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38990](https://github.com/BerriAI/litellm/pull/38990)
- test(ui): budget DOM-structure assertions in dashboard tests by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39082](https://github.com/BerriAI/litellm/pull/39082)
- test(ui): assert DataTable behavior instead of DOM structure by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39084](https://github.com/BerriAI/litellm/pull/39084)
- test(ui): query the screen instead of the render result by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39085](https://github.com/BerriAI/litellm/pull/39085)
- fix(ui): stop checkboxes stretching to the full width of a form field by [@&#8203;yatishgoel](https://github.com/yatishgoel) in [#&#8203;39108](https://github.com/BerriAI/litellm/pull/39108)
- chore: bump litellm-enterprise 0.1.62 -> 0.1.63, litellm-proxy-extras 0.4.91 -> 0.4.92, litellm 1.100.0 -> 1.101.0 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39140](https://github.com/BerriAI/litellm/pull/39140)
- revert: restore search tool fallback when no router is configured by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39146](https://github.com/BerriAI/litellm/pull/39146)
- test(websearch): register configured search tool in pre-request hook test by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39074](https://github.com/BerriAI/litellm/pull/39074)
- feat(proxy): default to the v2 migration resolver, keep v1 as an opt-out by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;31125](https://github.com/BerriAI/litellm/pull/31125)
- build(deps): bump browserslist to 4.28.8 to clear osv-scan by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39142](https://github.com/BerriAI/litellm/pull/39142)
- fix(ui): render the skill detail page with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39130](https://github.com/BerriAI/litellm/pull/39130)
- feat: add Azure AI DeepSeek V4 Flash 0731 pricing by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39023](https://github.com/BerriAI/litellm/pull/39023)
- fix(streaming): keep response id stable across streamed chunks by [@&#8203;Timik232](https://github.com/Timik232) in [#&#8203;38106](https://github.com/BerriAI/litellm/pull/38106)
- test(e2e/ui): cover the Budgets page create, edit and delete flows by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39052](https://github.com/BerriAI/litellm/pull/39052)
- feat(dashscope): add QwenCloud and Qwen AI Platform provider aliases by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39149](https://github.com/BerriAI/litellm/pull/39149)
- fix(bedrock): forward native structured outputs on Invoke instead of silently inlining the schema by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39070](https://github.com/BerriAI/litellm/pull/39070)
- test(e2e/ui): cover creating, testing and deleting a guardrail by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39053](https://github.com/BerriAI/litellm/pull/39053)
- refactor(types): replace Any with precise types across 73 modules by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39104](https://github.com/BerriAI/litellm/pull/39104)
- feat(models): add Claude Fable 5.1 across Anthropic, Bedrock, Vertex AI, and Azure AI by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39148](https://github.com/BerriAI/litellm/pull/39148)
- feat(guardrails): add Alice guardrail by [@&#8203;seanyasno-af](https://github.com/seanyasno-af) in [#&#8203;38898](https://github.com/BerriAI/litellm/pull/38898)
- test(e2e/ui): cover the Logs page filter drawer by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39056](https://github.com/BerriAI/litellm/pull/39056)
- test(e2e/ui): stop the suite failing on things that are not regressions by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39063](https://github.com/BerriAI/litellm/pull/39063)
- test(e2e/ui): cover the team Settings tab by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39058](https://github.com/BerriAI/litellm/pull/39058)
- test(e2e/ui): cover the Usage page activity tabs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39061](https://github.com/BerriAI/litellm/pull/39061)
- fix(ui): render the guardrail garden detail page with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39131](https://github.com/BerriAI/litellm/pull/39131)
- fix(responses): tool call id shape breaks gpt-5 -> claude fallback conversations by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39144](https://github.com/BerriAI/litellm/pull/39144)
- fix(openai): drop tool\_choice when request has no tools on chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39147](https://github.com/BerriAI/litellm/pull/39147)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39141](https://github.com/BerriAI/litellm/pull/39141)
- test(ui): pick select options by role instead of by text by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39175](https://github.com/BerriAI/litellm/pull/39175)
- feat(cost): support time-based off-peak pricing in cost calculation by [@&#8203;Srivatsa03](https://github.com/Srivatsa03) in [#&#8203;31725](https://github.com/BerriAI/litellm/pull/31725)
- fix(openai): flatten top-level tool schema combinators on chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38839](https://github.com/BerriAI/litellm/pull/38839)
- fix(s3): bound s3 object keys and download filenames for long Responses API ids by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39164](https://github.com/BerriAI/litellm/pull/39164)
- revert: default the proxy back to the v1 migration resolver by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39178](https://github.com/BerriAI/litellm/pull/39178)
- fix(prometheus): bound requested\_model label cardinality on client failure paths by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39136](https://github.com/BerriAI/litellm/pull/39136)
- feat(ui): add search to the Agent Hub tab and admin agents table by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39155](https://github.com/BerriAI/litellm/pull/39155)
- fix(anthropic): fix response\_format for claude-fable-5-1 on Vertex AI and Bedrock by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39184](https://github.com/BerriAI/litellm/pull/39184)
- fix: keep litellm\_credential\_name from LiteLLM Params JSON and gate stored credential attach to proxy admins by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39047](https://github.com/BerriAI/litellm/pull/39047)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39186](https://github.com/BerriAI/litellm/pull/39186)
- test: exempt MockTransport request-shape embedding tests from VCR replay by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39185](https://github.com/BerriAI/litellm/pull/39185)
- fix(ui): render the logs Tools panel with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39129](https://github.com/BerriAI/litellm/pull/39129)
- fix(proxy): default max\_idle\_connection\_lifetime to 60s on DB URLs by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39134](https://github.com/BerriAI/litellm/pull/39134)
- fix(mcp): follow tools/list pagination from upstream servers by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39172](https://github.com/BerriAI/litellm/pull/39172)
- fix(proxy): resolve router model aliases in /utils/supported\_openai\_params by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39000](https://github.com/BerriAI/litellm/pull/39000)
- fix(azure): flatten top-level tool schema combinators on Azure chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38870](https://github.com/BerriAI/litellm/pull/38870)
- fix(bedrock): route streamed responses-API output through the unified guardrail by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38734](https://github.com/BerriAI/litellm/pull/38734)
- fix(ui): hide model write affordances from view-only admin sessions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38872](https://github.com/BerriAI/litellm/pull/38872)
- fix(cli): quote the Claude Code apiKeyHelper for cmd.exe on Windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39174](https://github.com/BerriAI/litellm/pull/39174)
- fix(logging): guarantee max\_parallel\_requests slot release when streaming logging fails by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39093](https://github.com/BerriAI/litellm/pull/39093)
- feat(alerting): slack alerts for per-user daily/monthly spend thresholds and spend anomaly detection by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38438](https://github.com/BerriAI/litellm/pull/38438)
- fix(docker): add public Wolfi apk repo to runtime image by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39033](https://github.com/BerriAI/litellm/pull/39033)
- fix(router): keep order fallback on the requested order level by [@&#8203;emerzon](https://github.com/emerzon) in [#&#8203;38969](https://github.com/BerriAI/litellm/pull/38969)
- test(e2e): cover retry-on-timeout and the context-window fallback by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39197](https://github.com/BerriAI/litellm/pull/39197)
- fix(budget): reject known estimates over remaining budget under fail\_closed\_budget\_enforcement by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39214](https://github.com/BerriAI/litellm/pull/39214)
- fix: stop a cleared Team field from blocking personal key creation by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39206](https://github.com/BerriAI/litellm/pull/39206)
- test: record each e2e test's source location in the JUnit report by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39209](https://github.com/BerriAI/litellm/pull/39209)
- feat(router): fall back on anthropic safeguard refusals on /v1/messages by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39157](https://github.com/BerriAI/litellm/pull/39157)
- fix(proxy): report requested model on Anthropic streaming message\_start by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35816](https://github.com/BerriAI/litellm/pull/35816)
- fix(helm): reuse the generated master key Secret on helm upgrade by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39219](https://github.com/BerriAI/litellm/pull/39219)
- fix(mcp): report per-server outcomes in aggregate REST tools/list by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39232](https://github.com/BerriAI/litellm/pull/39232)
- fix(cost-map): retry transient boot fetch failures and recover config deployments dropped by a stale cost map by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39230](https://github.com/BerriAI/litellm/pull/39230)
- perf(scim): resolve group members with one user table read per member by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39228](https://github.com/BerriAI/litellm/pull/39228)
- fix(docker): install bedrock-realtime extra in monolith proxy images by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39223](https://github.com/BerriAI/litellm/pull/39223)
- fix(aiohttp\_transport): map transport-internal CancelledError to a retryable ConnectError by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39240](https://github.com/BerriAI/litellm/pull/39240)
- fix(bedrock): gate Converse cachePoint emission on model prompt caching support by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39210](https://github.com/BerriAI/litellm/pull/39210)
- fix(datadog\_llm\_obs): send tool calls, tool results and cache tokens in DD's own fields by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39222](https://github.com/BerriAI/litellm/pull/39222)
- feat(prometheus): expose per-key and per-team rate limit allowed and used gauges by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39236](https://github.com/BerriAI/litellm/pull/39236)
- feat(scim): add placeholder listing and merge so a shadowed account can be healed by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39231](https://github.com/BerriAI/litellm/pull/39231)
- fix: normalize provider-specific cache token fields in OTel v2 usage by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39202](https://github.com/BerriAI/litellm/pull/39202)
- fix: stop deployment default API key limits leaking into provider requests by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39211](https://github.com/BerriAI/litellm/pull/39211)
- fix(proxy): keep passthrough logging metadata and model\_info dicts when team callbacks are wired by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39216](https://github.com/BerriAI/litellm/pull/39216)
- fix(guardrails): deliver modify\_response block as valid SSE on streaming chat and Responses by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39036](https://github.com/BerriAI/litellm/pull/39036)
- fix(search): forward search-tool params through the router, complete Parallel AI v1 param mapping by [@&#8203;jliounis](https://github.com/jliounis) in [#&#8203;37883](https://github.com/BerriAI/litellm/pull/37883)
- fix(bedrock): stop Converse crashing on bearer-token auth without SigV4 credentials by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39166](https://github.com/BerriAI/litellm/pull/39166)
- fix(docker): install saml extra in litellm-backend image by [@&#8203;ojensen-berri](https://github.com/ojensen-berri) in [#&#8203;39291](https://github.com/BerriAI/litellm/pull/39291)
- fix(guardrails): run apply\_guardrail-only providers in logging\_only mode by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39297](https://github.com/BerriAI/litellm/pull/39297)
- feat(gemini): day-0 pricing for gemini-3.8-flash by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39340](https://github.com/BerriAI/litellm/pull/39340)
- fix(vertex): avoid duplicate DeepSeek OCR model namespace by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39194](https://github.com/BerriAI/litellm/pull/39194)
- feat(streaming): carry final response cost on streamed usage by default by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39069](https://github.com/BerriAI/litellm/pull/39069)
- fix(rerank): map provider errors with the resolved provider on sync and async paths by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39176](https://github.com/BerriAI/litellm/pull/39176)
- test(e2e): read JUnit properties off the real collected pytest Item by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39246](https://github.com/BerriAI/litellm/pull/39246)
- feat(proxy): configurable display\_name for the Anthropic-shaped /v1/models listing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39238](https://github.com/BerriAI/litellm/pull/39238)
- fix(helm): scale the classic chart's HPA out at the documented 60 percent CPU by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35975](https://github.com/BerriAI/litellm/pull/35975)
- fix(gemini): return enabled thinking content by default by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39160](https://github.com/BerriAI/litellm/pull/39160)
- fix: run access group key sync UPDATEs on the writer, not the read replica by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39128](https://github.com/BerriAI/litellm/pull/39128)
- fix(models): registry audit 2026-09-01: openai realtime and long-context tiers, mistral aliases, voyage, xai, fireworks, together, scaleway, azure ai, govcloud, azure gov, cloudflare whisper, deprecation dates by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39170](https://github.com/BerriAI/litellm/pull/39170)
- fix: apply optional\_pre\_call\_checks and reject unsupported router settings on /config/update by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39249](https://github.com/BerriAI/litellm/pull/39249)
- fix(vector\_stores): s3 vectors search router bypass + rag query config drop + ui error swallow by [@&#8203;michelligabriele](https://github.com/michelligabriele) in [#&#8203;34788](https://github.com/BerriAI/litellm/pull/34788)
- fix(models): key Azure DeepSeek V4 Flash 0731 by its Foundry catalog id by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39341](https://github.com/BerriAI/litellm/pull/39341)
- fix(deps): raise the tornado and pypdf floors for six new advisories by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39188](https://github.com/BerriAI/litellm/pull/39188)
- fix(headroom): stop re-compressing retrieved CCR content in client tool loops by [@&#8203;QuantumBreakz](https://github.com/QuantumBreakz) in [#&#8203;38591](https://github.com/BerriAI/litellm/pull/38591)
- feat(agentcore-a2a): derive runtime session id from A2A message.contextId by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39371](https://github.com/BerriAI/litellm/pull/39371)
- fix(proxy): share per-model budget counters across replicas through the spend counter cache by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39375](https://github.com/BerriAI/litellm/pull/39375)
- fix(proxy-extras): give prisma migrate deploy its own timeout budget by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39365](https://github.com/BerriAI/litellm/pull/39365)
- fix(proxy): route container create and list through model\_list deployments by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39220](https://github.com/BerriAI/litellm/pull/39220)
- test(build): validate release wheel contracts by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39021](https://github.com/BerriAI/litellm/pull/39021)
- refactor(rust): extract domain-neutral Python interop by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39026](https://github.com/BerriAI/litellm/pull/39026)
- refactor(rust): standardize the core Error type by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39331](https://github.com/BerriAI/litellm/pull/39331)
- fix(ui): preserve full AgentCore runtime ARN in agent edit form by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39382](https://github.com/BerriAI/litellm/pull/39382)
- feat(ui): update OpenAI preset model tiers by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39396](https://github.com/BerriAI/litellm/pull/39396)
- fix(router): resolve realtime session model to routed deployment by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36811](https://github.com/BerriAI/litellm/pull/36811)
- fix(security): restrict and validate file uploads at /v1/files and /upload/logo by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39379](https://github.com/BerriAI/litellm/pull/39379)
- feat(auth): enforce configurable password policy and SSO-only login by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39381](https://github.com/BerriAI/litellm/pull/39381)
- fix(agents): redact secret litellm\_params fields from all /v1/agents responses by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39389](https://github.com/BerriAI/litellm/pull/39389)
- fix(otel): stamp Langfuse root observation input and output from the request task by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39369](https://github.com/BerriAI/litellm/pull/39369)
- fix(guardrails): track and tear down presidio sibling callbacks on delete and update by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39271](https://github.com/BerriAI/litellm/pull/39271)
- fix(spend): keep every-deployment scope on gateway cache-injection marks by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39241](https://github.com/BerriAI/litellm/pull/39241)
- fix(proxy/db): keep prisma predicates from raising TypeError under a mocked prisma module by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39253](https://github.com/BerriAI/litellm/pull/39253)
- fix(proxy): word database 503s by whether the fault is transient by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39256](https://github.com/BerriAI/litellm/pull/39256)
- refactor(utils): remove the dead get\_api\_key provider-key resolver by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39260](https://github.com/BerriAI/litellm/pull/39260)
- feat(mcp): semantic tool search for the native MCP Gateway by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39404](https://github.com/BerriAI/litellm/pull/39404)
- fix(logging): redact credential query params from the uvicorn access log by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39293](https://github.com/BerriAI/litellm/pull/39293)
- feat(model\_prices): add meta/muse-spark-1.3 and its contributor tier by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39417](https://github.com/BerriAI/litellm/pull/39417)
- refactor(core): move audio transcription into core by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39126](https://github.com/BerriAI/litellm/pull/39126)
- fix(proxy): build coordination Redis from REDIS\_\* env vars unconditionally by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39410](https://github.com/BerriAI/litellm/pull/39410)
- test: add interactive Rust Python parity harness by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39419](https://github.com/BerriAI/litellm/pull/39419)
- test(proxy): verify NO\_DOCS/NO\_REDOC/NO\_OPENAPI restrict every doc surface by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39378](https://github.com/BerriAI/litellm/pull/39378)
- test(bedrock): accept the router kwarg in the knowledge base search fake by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39420](https://github.com/BerriAI/litellm/pull/39420)
- refactor(python-bridge): split routes and add shared function tracing by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39031](https://github.com/BerriAI/litellm/pull/39031)
- fix(python-bridge): harden sync and async execution boundaries by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39332](https://github.com/BerriAI/litellm/pull/39332)
- refactor(python-bridge): declare sync and async routes once by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39333](https://github.com/BerriAI/litellm/pull/39333)
- feat(python): unify Rust opt-in and bridge policy by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39334](https://github.com/BerriAI/litellm/pull/39334)
- feat(router): add heuristic v2 complexity routing by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39276](https://github.com/BerriAI/litellm/pull/39276)
- fix(anthropic): upgrade legacy thinking to adaptive on adaptive-only Claude models for chat, Bedrock Converse, Invoke, Vertex AI, and Databricks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39159](https://github.com/BerriAI/litellm/pull/39159)
- fix(proxy): mark session/SSO/SAML cookies Secure behind a TLS-terminating reverse proxy by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39391](https://github.com/BerriAI/litellm/pull/39391)
- fix(bedrock): honor BEDROCK\_MANTLE\_API\_BASE on bedrock/mantle messages and chat URLs by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39364](https://github.com/BerriAI/litellm/pull/39364)
- fix(bedrock): strip client\_metadata from converse additionalModelRequestFields by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35967](https://github.com/BerriAI/litellm/pull/35967)
- chore(techdebt): clear fresh debt from the 2026-08-31 and 2026-09-01 windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39091](https://github.com/BerriAI/litellm/pull/39091)
- fix(mcp): cap tools preview and test-connection at the listing timeout and name the unreachable upstream by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38791](https://github.com/BerriAI/litellm/pull/38791)
- fix(hosted\_vllm): forward truncate\_prompt\_tokens on rerank requests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39363](https://github.com/BerriAI/litellm/pull/39363)
- fix(messages): drop cache\_control ttl on non-Anthropic /v1/messages passthrough by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39355](https://github.com/BerriAI/litellm/pull/39355)
- fix(bedrock\_mantle): carry per-request AWS credentials into chat completions SigV4 signing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39362](https://github.com/BerriAI/litellm/pull/39362)
- feat(router): add a hybrid classifier that defers near tier boundaries by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39403](https://github.com/BerriAI/litellm/pull/39403)
- fix: recover the v2 migration resolver from concurrent migrate deploy deadlocks by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39187](https://github.com/BerriAI/litellm/pull/39187)
- fix(ollama\_chat): stamp finish\_reason tool\_calls when tool calls streamed before the done chunk by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39010](https://github.com/BerriAI/litellm/pull/39010)
- fix(router): route Claude Code subagents through session router by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39239](https://github.com/BerriAI/litellm/pull/39239)
- fix(responses): keep namespace tools intact when a guardrail returns them unchanged by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39366](https://github.com/BerriAI/litellm/pull/39366)
- fix(vector-store): resolve embedding credentials per request by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;38936](https://github.com/BerriAI/litellm/pull/38936)
- test(e2e/ui): give the seeded users passwords that pass the default password policy by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39442](https://github.com/BerriAI/litellm/pull/39442)
- fix(http\_handler): honor HTTP(S)\_PROXY / NO\_PROXY when force\_ipv4 uses the httpx transport by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39443](https://github.com/BerriAI/litellm/pull/39443)
- fix(proxy): stop leaking internal exception details to clients by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39380](https://github.com/BerriAI/litellm/pull/39380)
- fix(guardrails): forward mode and streaming params to crowdstrike\_aidr handler by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39317](https://github.com/BerriAI/litellm/pull/39317)
- fix(mcp): gate the connect-time OBO pre-flight on the key's allowed servers by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39447](https://github.com/BerriAI/litellm/pull/39447)
- fix(responses): keep provider response headers in streaming logging callbacks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38131](https://github.com/BerriAI/litellm/pull/38131)
- fix(mcp): fence an outbound-token write against an overlapping invalidation by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35398](https://github.com/BerriAI/litellm/pull/35398)
- feat(cli): pre-fill the SSO verification code in the browser when the proxy allows it by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39428](https://github.com/BerriAI/litellm/pull/39428)
- fix(ui): paginate request logs by session groups server-side by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39257](https://github.com/BerriAI/litellm/pull/39257)
- feat(proxy): serve the auto-router preset catalog at runtime by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39412](https://github.com/BerriAI/litellm/pull/39412)
- docs: define Rust Python harness structure by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39456](https://github.com/BerriAI/litellm/pull/39456)
- fix(guardrails): apply PUT /guardrails/{id} to the serving worker immediately and reject invalid configs with 422 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38877](https://github.com/BerriAI/litellm/pull/38877)
- test(responses): expect the 404 OpenAI now returns for an unknown model by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39457](https://github.com/BerriAI/litellm/pull/39457)
- fix(guardrails): skip streaming guardrail rounds that re-scan cleared output by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39386](https://github.com/BerriAI/litellm/pull/39386)
- fix: keep litellm importable on Python 3.10 and guard 3.11-only typing imports in CI by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39448](https://github.com/BerriAI/litellm/pull/39448)
- fix(proxy): keep SpendLogs and callback session ids in sync when the request has none by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39450](https://github.com/BerriAI/litellm/pull/39450)
- feat(router): arm safeguard-refusal fallback on generic chains when no content-policy list exists by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39274](https://github.com/BerriAI/litellm/pull/39274)
- feat(azure): support credential chain for storage by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39229](https://github.com/BerriAI/litellm/pull/39229)
- chore(crowdstrike): expect the deduped end-of-stream scan in crowdstrike cadence test by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39467](https://github.com/BerriAI/litellm/pull/39467)
- fix(model\_armor): handle Anthropic Messages and Responses streams in post\_call by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39181](https://github.com/BerriAI/litellm/pull/39181)
- test: add OCR python-to-rust test parity ledger (WIP) by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39434](https://github.com/BerriAI/litellm/pull/39434)
- feat(complexity\_router): opt-in modality override of a kept session-affinity pin by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39454](https://github.com/BerriAI/litellm/pull/39454)
- feat(datadog\_llm\_obs): cost tag dimensions, router decision fields, reasoning token metric, redaction gating by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39402](https://github.com/BerriAI/litellm/pull/39402)
- test(rust-python-harness): wire existing e2e SDK tests into the matrix by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39463](https://github.com/BerriAI/litellm/pull/39463)
- fix(mcp): never exchange the LiteLLM virtual key as the upstream subject token by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39446](https://github.com/BerriAI/litellm/pull/39446)
- test: add mistral ocr transformation parity coverage by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39482](https://github.com/BerriAI/litellm/pull/39482)
- test(vector-store): accept embedding\_executor in the Bedrock KB hook fake handler by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39472](https://github.com/BerriAI/litellm/pull/39472)
- refactor(s3\_vectors): embed search queries through the shared vector store executor by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39474](https://github.com/BerriAI/litellm/pull/39474)
- fix(xai): bill from the cost xAI reports instead of recomputing it (internal copy of [#&#8203;36281](https://github.com/BerriAI/litellm/issues/36281)) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39441](https://github.com/BerriAI/litellm/pull/39441)
- feat(ui): add 1M context auto-router preset by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39490](https://github.com/BerriAI/litellm/pull/39490)
- fix(ui): stop the create team form resetting organization and models by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39476](https://github.com/BerriAI/litellm/pull/39476)
- fix(ui): read the preset catalog at runtime in the dashboard tests by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39478](https://github.com/BerriAI/litellm/pull/39478)
- fix(sso): resolve multi-valued role claims to the highest privilege role by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39480](https://github.com/BerriAI/litellm/pull/39480)
- fix(guardrail): hide-secrets playground redaction and guardrail telemetry by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39398](https://github.com/BerriAI/litellm/pull/39398)
- fix(test): drop the duplicate embedding\_executor arg in the Bedrock KB fake handler by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39502](https://github.com/BerriAI/litellm/pull/39502)
- fix(ui): keep Virtual Keys list state in the URL so it survives leaving the page by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39481](https://github.com/BerriAI/litellm/pull/39481)
- fix(proxy): 404 a credential delete that matched nothing, and raise instead of return by [@&#8203;eeshsaxena](https://github.com/eeshsaxena) in [#&#8203;36260](https://github.com/BerriAI/litellm/pull/36260)
- fix(proxy-extras): only spend a migrate-deploy attempt when a pass made no progress by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39506](https://github.com/BerriAI/litellm/pull/39506)
- feat(cli): enable Claude Code gateway model discovery by default in lite claude by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39445](https://github.com/BerriAI/litellm/pull/39445)
- fix(docker): bump nginx runtime to 1.31.5-alpine3.24 and pin digest by [@&#8203;rakeshrepository](https://github.com/rakeshrepository) in [#&#8203;39561](https://github.com/BerriAI/litellm/pull/39561)
- fix: 1.99.0-rc2 UI bug batch (empty org on key create, session pagination, access group rename/delete) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39436](https://github.com/BerriAI/litellm/pull/39436)
- feat(auto-router): support classifier reasoning effort by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39372](https://github.com/BerriAI/litellm/pull/39372)
- fix(ui): replace the key detail URL entry when a virtual key is rotated by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39471](https://github.com/BerriAI/litellm/pull/39471)
- test(timeout): time out against the local fake endpoint instead of api.openai.com by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39583](https://github.com/BerriAI/litellm/pull/39583)
- test(harness): add OCR parity with migration strategy runners by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;38765](https://github.com/BerriAI/litellm/pull/38765)
- fix(databricks): strip thinking\_blocks and reasoning\_content from outbound messages by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39409](https://github.com/BerriAI/litellm/pull/39409)
- test(ocr): record provider fixtures in the migration harness by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39425](https://github.com/BerriAI/litellm/pull/39425)
- feat(ui): keyset-paginate request logs by session trace by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38794](https://github.com/BerriAI/litellm/pull/38794)
- fix(proxy/db): translate libpq sslrootcert and verify-\* into Prisma's strict TLS params by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39563](https://github.com/BerriAI/litellm/pull/39563)
- fix(agents): keep the published agent in public\_agent\_groups by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39554](https://github.com/BerriAI/litellm/pull/39554)
- fix(mcp): scope allow-all servers to virtual keys by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39531](https://github.com/BerriAI/litellm/pull/39531)
- fix(team): generate team IDs for blank input by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39571](https://github.com/BerriAI/litellm/pull/39571)
- fix(bedrock\_mantle): stop dropping the web\_search tool on /v1/responses by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35987](https://github.com/BerriAI/litellm/pull/35987)
- chore: bump litellm-enterprise 0.1.63 -> 0.1.64, litellm-proxy-extras 0.4.92 -> 0.4.93 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39595](https://github.com/BerriAI/litellm/pull/39595)
- fix(images): forward gpt-image supported params like background to OpenAI and Azure by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39525](https://github.com/BerriAI/litellm/pull/39525)
- fix(proxy): return persisted team memberships from /user/new so first CLI login gets the default team by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39545](https://github.com/BerriAI/litellm/pull/39545)
- fix(spend\_tracking): add missing\_session\_id: omit to leave SpendLogs.session\_id null without a client session by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39458](https://github.com/BerriAI/litellm/pull/39458)
- fix: stop a cleared Organization field from failing key creation by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39316](https://github.com/BerriAI/litellm/pull/39316)
- fix(ui): show MCP servers and agents inherited from access groups on team overview by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39215](https://github.com/BerriAI/litellm/pull/39215)
- fix(proxy): expose configured mode for auto-router models by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39619](https://github.com/BerriAI/litellm/pull/39619)
- fix(ui): aggregate session token usage in the logs table by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39598](https://github.com/BerriAI/litellm/pull/39598)
- fix(cost): apply off\_peak\_pricing in the dashscope cost calculator by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39592](https://github.com/BerriAI/litellm/pull/39592)
- test(bedrock): drop EOL cohere.command-r-plus-v1:0 from local\_testing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39608](https://github.com/BerriAI/litellm/pull/39608)
- fix(openai): default stream usage on PrivateLink and regional api.openai.com hosts by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39614](https://github.com/BerriAI/litellm/pull/39614)
- fix(proxy): drop anthropic-beta on the Vertex passthrough count-tokens route by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39597](https://github.com/BerriAI/litellm/pull/39597)
- fix(headroom): resolve CCR retrieval on streaming /v1/responses by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38808](https://github.com/BerriAI/litellm/pull/38808)
- fix(openai): bridge gpt-5.4+ tool calls to /v1/responses on every api.openai.com host by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39587](https://github.com/BerriAI/litellm/pull/39587)
- fix(router): pin JWT-authenticated callers by user id in deployment\_affinity by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39594](https://github.com/BerriAI/litellm/pull/39594)
- fix(cost): bill bedrock\_mantle web search at $12 per 1k queries using Bedrock's reported count by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39610](https://github.com/BerriAI/litellm/pull/39610)
- fix(azure\_ai): don't reclassify Foundry deployments as azure provider by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38975](https://github.com/BerriAI/litellm/pull/38975)
- fix(vector\_stores): only list vector stores the caller was granted by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39612](https://github.com/BerriAI/litellm/pull/39612)
- feat(models): add gpt-6-astra pricing and metadata by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39622](https://github.com/BerriAI/litellm/pull/39622)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39593](https://github.com/BerriAI/litellm/pull/39593)
- feat(router): limit heuristic\_v2 auto-routers to one without the auto\_router license feature by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39468](https://github.com/BerriAI/litellm/pull/39468)
- fix(ui): clear agents when updating team permissions by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39600](https://github.com/BerriAI/litellm/pull/39600)
- fix(auto\_router): bill the routing embedding to the caller's key and team by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39532](https://github.com/BerriAI/litellm/pull/39532)
- test(router): cover get\_configured\_mode so router\_code\_coverage passes by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;39630](https://github.com/BerriAI/litellm/pull/39630)
- fix: treat gpt-6 names as the gpt-5 request family in OpenAI and Azure configs by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39631](https://github.com/BerriAI/litellm/pull/39631)
- fix(prompts): key the in-memory prompt registry by environment by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38440](https://github.com/BerriAI/litellm/pull/38440)
- fix(ui): let the Internal Users search box match user\_id as well as email by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39604](https://github.com/BerriAI/litellm/pull/39604)
- test(responses): bound the background stream cancel e2e so an upstream stall skips fast by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39617](https://github.com/BerriAI/litellm/pull/39617)
- fix(vertex): add the API version to versionless project routes on the Vertex passthrough by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39625](https://github.com/BerriAI/litellm/pull/39625)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39648](https://github.com/BerriAI/litellm/pull/39648)
- fix(spend\_tracking): key /v1/messages spend rows on the msg\_ id the client received by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39511](https://github.com/BerriAI/litellm/pull/39511)
- ci(rust): build and test the ai-gateway server feature by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39493](https://github.com/BerriAI/litellm/pull/39493)
- ci(ui): run the UI build check through the image's ui-builder stage by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39496](https://github.com/BerriAI/litellm/pull/39496)
- fix(proxy): parse numeric multipart fields on /v1/images/edits back into numbers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39510](https://github.com/BerriAI/litellm/pull/39510)
- fix(guardrails): remove the module-global translation mapping that leaked between tests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39543](https://github.com/BerriAI/litellm/pull/39543)
- feat(azure\_ai): add grok-4.6 to the model cost map by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39426](https://github.com/BerriAI/litellm/pull/39426)
- fix: attach vector store search\_results when a guardrail is registered by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38984](https://github.com/BerriAI/litellm/pull/38984)
- fix(proxy): stop putting the literal string "None" in error payloads by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39521](https://github.com/BerriAI/litellm/pull/39521)
- fix(router): keep retry breadcrumbs per request and out of the request snapshot by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39491](https://github.com/BerriAI/litellm/pull/39491)
- fix(vector-stores): survive a failing vector store search in the chat completions hook by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39495](https://github.com/BerriAI/litellm/pull/39495)
- fix(utils): redact credential kwargs from the set\_verbose request line by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39526](https://github.com/BerriAI/litellm/pull/39526)
- fix(bedrock): skip the SigV4 credential chain when a bearer token is configured by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39411](https://github.com/BerriAI/litellm/pull/39411)
- fix(proxy-extras): kill the whole Prisma process group when a command times out by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39466](https://github.com/BerriAI/litellm/pull/39466)
- fix(rag): forward the managed vector store's params to the search call by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39452](https://github.com/BerriAI/litellm/pull/39452)
- fix(utils): redact credentials nested in extra\_body on the verbose optional-params line by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39538](https://github.com/BerriAI/litellm/pull/39538)
- fix(cont…
GiorgioAresu pushed a commit to GiorgioAresu/home-ops that referenced this pull request Sep 18, 2026
…101.0) (#2106)

This PR contains the following updates:

| Package | Update | Change |
|---|---|---|
| [ghcr.io/berriai/litellm](https://images.chainguard.dev/directory/image/wolfi-base/overview) ([source](https://github.com/BerriAI/litellm)) | minor | `v1.100.1` → `v1.101.0` |

---

> ⚠️ **Warning**
>
> Some dependencies could not be looked up. Check the [Dependency Dashboard](issues/6) for more information.

---

### Release Notes

<details>
<summary>BerriAI/litellm (ghcr.io/berriai/litellm)</summary>

### [`v1.101.0`](https://github.com/BerriAI/litellm/releases/tag/v1.101.0)

[Compare Source](https://github.com/BerriAI/litellm/compare/v1.100.1...v1.101.0)

#### Verify Docker Image Signature

All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0).

**Verify using the pinned commit hash (recommended):**

A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \
  ghcr.io/berriai/litellm:v1.101.0
```

**Verify using the release tag (convenience):**

Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/v1.101.0/cosign.pub \
  ghcr.io/berriai/litellm:v1.101.0
```

Expected output:

```
The following checks were performed on each of these signatures:
  - The cosign claims were validated
  - The signatures were verified against the specified public key
```

***

#### What's Changed

- fix(proxy): emit timing headers and overhead for /v1/messages and /v1/responses by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38840](https://github.com/BerriAI/litellm/pull/38840)
- fix(tests): derive the no-cache-read-rate savings baseline from the model map by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38863](https://github.com/BerriAI/litellm/pull/38863)
- chore(typing): clear Any seams across 47 files, ratchet basedpyright ceilings -3,302 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;37778](https://github.com/BerriAI/litellm/pull/37778)
- chore(typing): clear 1.2k basedpyright Any errors across 16 hotspot files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36722](https://github.com/BerriAI/litellm/pull/36722)
- feat(bedrock): honor streaming buffer/sampling config for unbuffered post\_call scans by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38722](https://github.com/BerriAI/litellm/pull/38722)
- feat(cli): set ENABLE\_TOOL\_SEARCH=true for lite claude by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38942](https://github.com/BerriAI/litellm/pull/38942)
- fix(proxy): deliver budget alerts on webhook-only alerting and accept ALERTING\_WEBHOOK\_URL by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38441](https://github.com/BerriAI/litellm/pull/38441)
- docs(claude.md): require tests to check behavior, not code structure by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;38772](https://github.com/BerriAI/litellm/pull/38772)
- chore(newrelic): cover static default\_team\_settings per-team routing by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38857](https://github.com/BerriAI/litellm/pull/38857)
- fix: update stale source URLs and deprecation dates in model cost map by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38801](https://github.com/BerriAI/litellm/pull/38801)
- feat(ci): close duplicate issues after a 3-day grace period by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;38381](https://github.com/BerriAI/litellm/pull/38381)
- docs(proxy): clarify spend semantics on /v2/user/info and /user/daily/activity by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38883](https://github.com/BerriAI/litellm/pull/38883)
- fix(guardrails): configure Prompt Security file timeout policy by [@&#8203;davida-ps](https://github.com/davida-ps) in [#&#8203;38083](https://github.com/BerriAI/litellm/pull/38083)
- fix(bedrock): stop duplicating Converse config blocks inside inferenceConfig by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38993](https://github.com/BerriAI/litellm/pull/38993)
- fix(guardrails): exclude images from HiddenLayer v1 scans by [@&#8203;Ashton-Sidhu](https://github.com/Ashton-Sidhu) in [#&#8203;29210](https://github.com/BerriAI/litellm/pull/29210)
- feat(spend\_tracking): persist router metadata in spend logs for internal router models by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39001](https://github.com/BerriAI/litellm/pull/39001)
- fix(vertex\_ai): graft default vertex path when api\_base has a version-only path by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38986](https://github.com/BerriAI/litellm/pull/38986)
- fix(proxy): allow unblocking customers via /customer/update by [@&#8203;cat0825](https://github.com/cat0825) in [#&#8203;34696](https://github.com/BerriAI/litellm/pull/34696)
- feat(openai): support workload identity federation (OIDC token exchange) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38995](https://github.com/BerriAI/litellm/pull/38995)
- fix(otel): emit cache token counts on OTel v2 LLM spans by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38716](https://github.com/BerriAI/litellm/pull/38716)
- feat(proxy): add /v1/responses/input\_tokens token counting endpoint by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38997](https://github.com/BerriAI/litellm/pull/38997)
- fix(docker): bump wolfi-base for glibc 2.44 and pin apk python to 3.13 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38917](https://github.com/BerriAI/litellm/pull/38917)
- fix(docker): bump wolfi-base for glibc 2.44 and pin apk python to 3.13 in migrations image by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38973](https://github.com/BerriAI/litellm/pull/38973)
- feat(friendli): add zai-org/GLM-5.3-Flash model pricing by [@&#8203;Lee-Si-Yoon](https://github.com/Lee-Si-Yoon) in [#&#8203;38880](https://github.com/BerriAI/litellm/pull/38880)
- chore(techdebt): clear fresh debt from the 2026-08-29 and 2026-08-30 windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38884](https://github.com/BerriAI/litellm/pull/38884)
- fix(bedrock): surface Nova Sonic user transcripts, speech events, and usage in realtime API by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38597](https://github.com/BerriAI/litellm/pull/38597)
- fix(guardrails): carry Anthropic url image sources through to guardrails by [@&#8203;samtsai15](https://github.com/samtsai15) in [#&#8203;38940](https://github.com/BerriAI/litellm/pull/38940)
- feat(friendli): add zai-org/GLM-5.3 model pricing by [@&#8203;Lee-Si-Yoon](https://github.com/Lee-Si-Yoon) in [#&#8203;38881](https://github.com/BerriAI/litellm/pull/38881)
- fix(router): apply model renames to the in-memory deployment list by [@&#8203;yatishgoel](https://github.com/yatishgoel) in [#&#8203;38479](https://github.com/BerriAI/litellm/pull/38479)
- test(e2e): cover SCIM token creation and SCIM API auth in the Admin UI suite by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39027](https://github.com/BerriAI/litellm/pull/39027)
- feat(gigachat): add native API passthrough routes with spend logging by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38913](https://github.com/BerriAI/litellm/pull/38913)
- feat(gigachat): add passthrough gigachat route by [@&#8203;KnyazSh](https://github.com/KnyazSh) in [#&#8203;25886](https://github.com/BerriAI/litellm/pull/25886)
- feat(complexity-router): add classification\_mode to skip classifier on continuation turns by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38861](https://github.com/BerriAI/litellm/pull/38861)
- fix(proxy): preserve model table columns on master key rotation by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38878](https://github.com/BerriAI/litellm/pull/38878)
- fix(speech): stop forwarding response\_format as a chat param for Gemini TTS by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38819](https://github.com/BerriAI/litellm/pull/38819)
- fix(proxy): return 200 from /model/block and /model/unblock instead of 500 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38873](https://github.com/BerriAI/litellm/pull/38873)
- feat(complexity\_router): escalate oversized prompts to a tier that fits before dispatch by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38844](https://github.com/BerriAI/litellm/pull/38844)
- feat(shadow\_eval): target teams and users so JWT-auth traffic can be evaluated by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39015](https://github.com/BerriAI/litellm/pull/39015)
- fix(anthropic\_messages): drain upstream in a detached pump so client … by [@&#8203;nuernber](https://github.com/nuernber) in [#&#8203;36008](https://github.com/BerriAI/litellm/pull/36008)
- refactor(proxy): bound the budget window seed by time instead of request ids by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;38851](https://github.com/BerriAI/litellm/pull/38851)
- fix(proxy): ship psycopg so partitioned SpendLogs detection actually runs by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;38994](https://github.com/BerriAI/litellm/pull/38994)
- test(e2e): assert user-observable behavior instead of DOM structure by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39016](https://github.com/BerriAI/litellm/pull/39016)
- build(rust): configure native extension profiles by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39020](https://github.com/BerriAI/litellm/pull/39020)
- fix(ui): keep litellm\_credential\_name from LiteLLM Params JSON when no credential is selected by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39005](https://github.com/BerriAI/litellm/pull/39005)
- Revert "fix(ui): keep litellm\_credential\_name from LiteLLM Params JSON when no credential is selected" by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39046](https://github.com/BerriAI/litellm/pull/39046)
- fix(auth): quiet malformed virtual key rejections to stdout by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38838](https://github.com/BerriAI/litellm/pull/38838)
- fix(proxy): wire team-level logging callbacks into passthrough endpoints by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38979](https://github.com/BerriAI/litellm/pull/38979)
- feat(complexity\_router): opt-in modality-based capability routing for image requests by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39032](https://github.com/BerriAI/litellm/pull/39032)
- fix(ui): let the auto-router scoring tier list follow the theme by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39040](https://github.com/BerriAI/litellm/pull/39040)
- feat(ui): auto-router controls for context-window escalation by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39054](https://github.com/BerriAI/litellm/pull/39054)
- fix(redis): coerce env var string types and fix param discovery through decorator wrappers by [@&#8203;koladefaj](https://github.com/koladefaj) in [#&#8203;30644](https://github.com/BerriAI/litellm/pull/30644)
- feat(key management): show budget window usage on /key/info by [@&#8203;Thijmen](https://github.com/Thijmen) in [#&#8203;37044](https://github.com/BerriAI/litellm/pull/37044)
- fix(websearch): reject invalid explicit search tool selections by [@&#8203;georgeatparallel](https://github.com/georgeatparallel) in [#&#8203;38113](https://github.com/BerriAI/litellm/pull/38113)
- feat(shadow\_eval): compare several auto-routers on one job's sampled traffic by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39028](https://github.com/BerriAI/litellm/pull/39028)
- fix(speech): honor pcm/wav response\_format for Gemini TTS and reject unsupported containers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38868](https://github.com/BerriAI/litellm/pull/38868)
- fix(proxy): match /v1/audio/speech content-type to the returned audio format by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38798](https://github.com/BerriAI/litellm/pull/38798)
- test(e2e): drop the two mgmt registry cells no shared-proxy test can cover by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39055](https://github.com/BerriAI/litellm/pull/39055)
- feat(ui): one classification frequency picker for complexity auto-routers by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39042](https://github.com/BerriAI/litellm/pull/39042)
- test(e2e/ui): automate 8 manual QA checklist flows by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39025](https://github.com/BerriAI/litellm/pull/39025)
- fix(key\_management): allow non-admin key\_type preset transitions on /key/update by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39051](https://github.com/BerriAI/litellm/pull/39051)
- chore(typing): clear 1.1k basedpyright Any errors across 53 backend files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38796](https://github.com/BerriAI/litellm/pull/38796)
- fix(openai): forward reasoning\_effort for unknown model aliases instead of failing closed by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39065](https://github.com/BerriAI/litellm/pull/39065)
- test(e2e-ui): poll credential availability before Test Connect to deflake multi-instance runs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39073](https://github.com/BerriAI/litellm/pull/39073)
- feat(ui): modality routing toggle on the auto-router create and edit forms by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39059](https://github.com/BerriAI/litellm/pull/39059)
- fix(proxy): include litellm\_model\_table in GET /v2/team/list by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39045](https://github.com/BerriAI/litellm/pull/39045)
- fix(bedrock): mask signed request headers in guardrail debug log by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39044](https://github.com/BerriAI/litellm/pull/39044)
- fix(bedrock): forward aws\_external\_id in files and batches credential loading by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39066](https://github.com/BerriAI/litellm/pull/39066)
- fix(mcp): persist alias MCP grants verbatim instead of rewriting to local server ids by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39119](https://github.com/BerriAI/litellm/pull/39119)
- fix(responses): json-encode object tool call arguments in the chat completions bridge by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35417](https://github.com/BerriAI/litellm/pull/35417)
- fix(cost): bill OCR annotation pages via annotation\_cost\_per\_page by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38985](https://github.com/BerriAI/litellm/pull/38985)
- fix(policy\_engine): restore request guardrails list after pipeline allow by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39038](https://github.com/BerriAI/litellm/pull/39038)
- fix(embeddings): omit encoding\_format when the client omits it on OpenAI-compatible calls by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38774](https://github.com/BerriAI/litellm/pull/38774)
- test: deflake MCP registry state, savings cost map, and MCP identity env reload tests by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38891](https://github.com/BerriAI/litellm/pull/38891)
- feat(helm): add Argo CD PreSync hook and rollout strategy knobs to the componentized chart by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39112](https://github.com/BerriAI/litellm/pull/39112)
- fix(registry): veo 3.1 pricing tiers + roll up open registry PRs (glm-5.2, Qwen3.8-Flash, gemma-4-31b, scribe\_v2, fireworks/databricks deepseek v4) + deprecation dates by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38990](https://github.com/BerriAI/litellm/pull/38990)
- test(ui): budget DOM-structure assertions in dashboard tests by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39082](https://github.com/BerriAI/litellm/pull/39082)
- test(ui): assert DataTable behavior instead of DOM structure by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39084](https://github.com/BerriAI/litellm/pull/39084)
- test(ui): query the screen instead of the render result by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39085](https://github.com/BerriAI/litellm/pull/39085)
- fix(ui): stop checkboxes stretching to the full width of a form field by [@&#8203;yatishgoel](https://github.com/yatishgoel) in [#&#8203;39108](https://github.com/BerriAI/litellm/pull/39108)
- chore: bump litellm-enterprise 0.1.62 -> 0.1.63, litellm-proxy-extras 0.4.91 -> 0.4.92, litellm 1.100.0 -> 1.101.0 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39140](https://github.com/BerriAI/litellm/pull/39140)
- revert: restore search tool fallback when no router is configured by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39146](https://github.com/BerriAI/litellm/pull/39146)
- test(websearch): register configured search tool in pre-request hook test by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39074](https://github.com/BerriAI/litellm/pull/39074)
- feat(proxy): default to the v2 migration resolver, keep v1 as an opt-out by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;31125](https://github.com/BerriAI/litellm/pull/31125)
- build(deps): bump browserslist to 4.28.8 to clear osv-scan by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39142](https://github.com/BerriAI/litellm/pull/39142)
- fix(ui): render the skill detail page with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39130](https://github.com/BerriAI/litellm/pull/39130)
- feat: add Azure AI DeepSeek V4 Flash 0731 pricing by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39023](https://github.com/BerriAI/litellm/pull/39023)
- fix(streaming): keep response id stable across streamed chunks by [@&#8203;Timik232](https://github.com/Timik232) in [#&#8203;38106](https://github.com/BerriAI/litellm/pull/38106)
- test(e2e/ui): cover the Budgets page create, edit and delete flows by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39052](https://github.com/BerriAI/litellm/pull/39052)
- feat(dashscope): add QwenCloud and Qwen AI Platform provider aliases by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39149](https://github.com/BerriAI/litellm/pull/39149)
- fix(bedrock): forward native structured outputs on Invoke instead of silently inlining the schema by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39070](https://github.com/BerriAI/litellm/pull/39070)
- test(e2e/ui): cover creating, testing and deleting a guardrail by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39053](https://github.com/BerriAI/litellm/pull/39053)
- refactor(types): replace Any with precise types across 73 modules by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39104](https://github.com/BerriAI/litellm/pull/39104)
- feat(models): add Claude Fable 5.1 across Anthropic, Bedrock, Vertex AI, and Azure AI by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39148](https://github.com/BerriAI/litellm/pull/39148)
- feat(guardrails): add Alice guardrail by [@&#8203;seanyasno-af](https://github.com/seanyasno-af) in [#&#8203;38898](https://github.com/BerriAI/litellm/pull/38898)
- test(e2e/ui): cover the Logs page filter drawer by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39056](https://github.com/BerriAI/litellm/pull/39056)
- test(e2e/ui): stop the suite failing on things that are not regressions by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39063](https://github.com/BerriAI/litellm/pull/39063)
- test(e2e/ui): cover the team Settings tab by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39058](https://github.com/BerriAI/litellm/pull/39058)
- test(e2e/ui): cover the Usage page activity tabs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39061](https://github.com/BerriAI/litellm/pull/39061)
- fix(ui): render the guardrail garden detail page with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39131](https://github.com/BerriAI/litellm/pull/39131)
- fix(responses): tool call id shape breaks gpt-5 -> claude fallback conversations by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39144](https://github.com/BerriAI/litellm/pull/39144)
- fix(openai): drop tool\_choice when request has no tools on chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39147](https://github.com/BerriAI/litellm/pull/39147)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39141](https://github.com/BerriAI/litellm/pull/39141)
- test(ui): pick select options by role instead of by text by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39175](https://github.com/BerriAI/litellm/pull/39175)
- feat(cost): support time-based off-peak pricing in cost calculation by [@&#8203;Srivatsa03](https://github.com/Srivatsa03) in [#&#8203;31725](https://github.com/BerriAI/litellm/pull/31725)
- fix(openai): flatten top-level tool schema combinators on chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38839](https://github.com/BerriAI/litellm/pull/38839)
- fix(s3): bound s3 object keys and download filenames for long Responses API ids by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39164](https://github.com/BerriAI/litellm/pull/39164)
- revert: default the proxy back to the v1 migration resolver by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39178](https://github.com/BerriAI/litellm/pull/39178)
- fix(prometheus): bound requested\_model label cardinality on client failure paths by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39136](https://github.com/BerriAI/litellm/pull/39136)
- feat(ui): add search to the Agent Hub tab and admin agents table by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39155](https://github.com/BerriAI/litellm/pull/39155)
- fix(anthropic): fix response\_format for claude-fable-5-1 on Vertex AI and Bedrock by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39184](https://github.com/BerriAI/litellm/pull/39184)
- fix: keep litellm\_credential\_name from LiteLLM Params JSON and gate stored credential attach to proxy admins by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39047](https://github.com/BerriAI/litellm/pull/39047)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39186](https://github.com/BerriAI/litellm/pull/39186)
- test: exempt MockTransport request-shape embedding tests from VCR replay by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39185](https://github.com/BerriAI/litellm/pull/39185)
- fix(ui): render the logs Tools panel with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39129](https://github.com/BerriAI/litellm/pull/39129)
- fix(proxy): default max\_idle\_connection\_lifetime to 60s on DB URLs by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39134](https://github.com/BerriAI/litellm/pull/39134)
- fix(mcp): follow tools/list pagination from upstream servers by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39172](https://github.com/BerriAI/litellm/pull/39172)
- fix(proxy): resolve router model aliases in /utils/supported\_openai\_params by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39000](https://github.com/BerriAI/litellm/pull/39000)
- fix(azure): flatten top-level tool schema combinators on Azure chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38870](https://github.com/BerriAI/litellm/pull/38870)
- fix(bedrock): route streamed responses-API output through the unified guardrail by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38734](https://github.com/BerriAI/litellm/pull/38734)
- fix(ui): hide model write affordances from view-only admin sessions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38872](https://github.com/BerriAI/litellm/pull/38872)
- fix(cli): quote the Claude Code apiKeyHelper for cmd.exe on Windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39174](https://github.com/BerriAI/litellm/pull/39174)
- fix(logging): guarantee max\_parallel\_requests slot release when streaming logging fails by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39093](https://github.com/BerriAI/litellm/pull/39093)
- feat(alerting): slack alerts for per-user daily/monthly spend thresholds and spend anomaly detection by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38438](https://github.com/BerriAI/litellm/pull/38438)
- fix(docker): add public Wolfi apk repo to runtime image by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39033](https://github.com/BerriAI/litellm/pull/39033)
- fix(router): keep order fallback on the requested order level by [@&#8203;emerzon](https://github.com/emerzon) in [#&#8203;38969](https://github.com/BerriAI/litellm/pull/38969)
- test(e2e): cover retry-on-timeout and the context-window fallback by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39197](https://github.com/BerriAI/litellm/pull/39197)
- fix(budget): reject known estimates over remaining budget under fail\_closed\_budget\_enforcement by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39214](https://github.com/BerriAI/litellm/pull/39214)
- fix: stop a cleared Team field from blocking personal key creation by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39206](https://github.com/BerriAI/litellm/pull/39206)
- test: record each e2e test's source location in the JUnit report by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39209](https://github.com/BerriAI/litellm/pull/39209)
- feat(router): fall back on anthropic safeguard refusals on /v1/messages by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39157](https://github.com/BerriAI/litellm/pull/39157)
- fix(proxy): report requested model on Anthropic streaming message\_start by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35816](https://github.com/BerriAI/litellm/pull/35816)
- fix(helm): reuse the generated master key Secret on helm upgrade by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39219](https://github.com/BerriAI/litellm/pull/39219)
- fix(mcp): report per-server outcomes in aggregate REST tools/list by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39232](https://github.com/BerriAI/litellm/pull/39232)
- fix(cost-map): retry transient boot fetch failures and recover config deployments dropped by a stale cost map by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39230](https://github.com/BerriAI/litellm/pull/39230)
- perf(scim): resolve group members with one user table read per member by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39228](https://github.com/BerriAI/litellm/pull/39228)
- fix(docker): install bedrock-realtime extra in monolith proxy images by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39223](https://github.com/BerriAI/litellm/pull/39223)
- fix(aiohttp\_transport): map transport-internal CancelledError to a retryable ConnectError by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39240](https://github.com/BerriAI/litellm/pull/39240)
- fix(bedrock): gate Converse cachePoint emission on model prompt caching support by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39210](https://github.com/BerriAI/litellm/pull/39210)
- fix(datadog\_llm\_obs): send tool calls, tool results and cache tokens in DD's own fields by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39222](https://github.com/BerriAI/litellm/pull/39222)
- feat(prometheus): expose per-key and per-team rate limit allowed and used gauges by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39236](https://github.com/BerriAI/litellm/pull/39236)
- feat(scim): add placeholder listing and merge so a shadowed account can be healed by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39231](https://github.com/BerriAI/litellm/pull/39231)
- fix: normalize provider-specific cache token fields in OTel v2 usage by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39202](https://github.com/BerriAI/litellm/pull/39202)
- fix: stop deployment default API key limits leaking into provider requests by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39211](https://github.com/BerriAI/litellm/pull/39211)
- fix(proxy): keep passthrough logging metadata and model\_info dicts when team callbacks are wired by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39216](https://github.com/BerriAI/litellm/pull/39216)
- fix(guardrails): deliver modify\_response block as valid SSE on streaming chat and Responses by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39036](https://github.com/BerriAI/litellm/pull/39036)
- fix(search): forward search-tool params through the router, complete Parallel AI v1 param mapping by [@&#8203;jliounis](https://github.com/jliounis) in [#&#8203;37883](https://github.com/BerriAI/litellm/pull/37883)
- fix(bedrock): stop Converse crashing on bearer-token auth without SigV4 credentials by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39166](https://github.com/BerriAI/litellm/pull/39166)
- fix(docker): install saml extra in litellm-backend image by [@&#8203;ojensen-berri](https://github.com/ojensen-berri) in [#&#8203;39291](https://github.com/BerriAI/litellm/pull/39291)
- fix(guardrails): run apply\_guardrail-only providers in logging\_only mode by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39297](https://github.com/BerriAI/litellm/pull/39297)
- feat(gemini): day-0 pricing for gemini-3.8-flash by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39340](https://github.com/BerriAI/litellm/pull/39340)
- fix(vertex): avoid duplicate DeepSeek OCR model namespace by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39194](https://github.com/BerriAI/litellm/pull/39194)
- feat(streaming): carry final response cost on streamed usage by default by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39069](https://github.com/BerriAI/litellm/pull/39069)
- fix(rerank): map provider errors with the resolved provider on sync and async paths by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39176](https://github.com/BerriAI/litellm/pull/39176)
- test(e2e): read JUnit properties off the real collected pytest Item by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39246](https://github.com/BerriAI/litellm/pull/39246)
- feat(proxy): configurable display\_name for the Anthropic-shaped /v1/models listing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39238](https://github.com/BerriAI/litellm/pull/39238)
- fix(helm): scale the classic chart's HPA out at the documented 60 percent CPU by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35975](https://github.com/BerriAI/litellm/pull/35975)
- fix(gemini): return enabled thinking content by default by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39160](https://github.com/BerriAI/litellm/pull/39160)
- fix: run access group key sync UPDATEs on the writer, not the read replica by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39128](https://github.com/BerriAI/litellm/pull/39128)
- fix(models): registry audit 2026-09-01: openai realtime and long-context tiers, mistral aliases, voyage, xai, fireworks, together, scaleway, azure ai, govcloud, azure gov, cloudflare whisper, deprecation dates by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39170](https://github.com/BerriAI/litellm/pull/39170)
- fix: apply optional\_pre\_call\_checks and reject unsupported router settings on /config/update by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39249](https://github.com/BerriAI/litellm/pull/39249)
- fix(vector\_stores): s3 vectors search router bypass + rag query config drop + ui error swallow by [@&#8203;michelligabriele](https://github.com/michelligabriele) in [#&#8203;34788](https://github.com/BerriAI/litellm/pull/34788)
- fix(models): key Azure DeepSeek V4 Flash 0731 by its Foundry catalog id by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39341](https://github.com/BerriAI/litellm/pull/39341)
- fix(deps): raise the tornado and pypdf floors for six new advisories by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39188](https://github.com/BerriAI/litellm/pull/39188)
- fix(headroom): stop re-compressing retrieved CCR content in client tool loops by [@&#8203;QuantumBreakz](https://github.com/QuantumBreakz) in [#&#8203;38591](https://github.com/BerriAI/litellm/pull/38591)
- feat(agentcore-a2a): derive runtime session id from A2A message.contextId by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39371](https://github.com/BerriAI/litellm/pull/39371)
- fix(proxy): share per-model budget counters across replicas through the spend counter cache by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39375](https://github.com/BerriAI/litellm/pull/39375)
- fix(proxy-extras): give prisma migrate deploy its own timeout budget by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39365](https://github.com/BerriAI/litellm/pull/39365)
- fix(proxy): route container create and list through model\_list deployments by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39220](https://github.com/BerriAI/litellm/pull/39220)
- test(build): validate release wheel contracts by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39021](https://github.com/BerriAI/litellm/pull/39021)
- refactor(rust): extract domain-neutral Python interop by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39026](https://github.com/BerriAI/litellm/pull/39026)
- refactor(rust): standardize the core Error type by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39331](https://github.com/BerriAI/litellm/pull/39331)
- fix(ui): preserve full AgentCore runtime ARN in agent edit form by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39382](https://github.com/BerriAI/litellm/pull/39382)
- feat(ui): update OpenAI preset model tiers by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39396](https://github.com/BerriAI/litellm/pull/39396)
- fix(router): resolve realtime session model to routed deployment by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36811](https://github.com/BerriAI/litellm/pull/36811)
- fix(security): restrict and validate file uploads at /v1/files and /upload/logo by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39379](https://github.com/BerriAI/litellm/pull/39379)
- feat(auth): enforce configurable password policy and SSO-only login by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39381](https://github.com/BerriAI/litellm/pull/39381)
- fix(agents): redact secret litellm\_params fields from all /v1/agents responses by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39389](https://github.com/BerriAI/litellm/pull/39389)
- fix(otel): stamp Langfuse root observation input and output from the request task by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39369](https://github.com/BerriAI/litellm/pull/39369)
- fix(guardrails): track and tear down presidio sibling callbacks on delete and update by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39271](https://github.com/BerriAI/litellm/pull/39271)
- fix(spend): keep every-deployment scope on gateway cache-injection marks by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39241](https://github.com/BerriAI/litellm/pull/39241)
- fix(proxy/db): keep prisma predicates from raising TypeError under a mocked prisma module by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39253](https://github.com/BerriAI/litellm/pull/39253)
- fix(proxy): word database 503s by whether the fault is transient by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39256](https://github.com/BerriAI/litellm/pull/39256)
- refactor(utils): remove the dead get\_api\_key provider-key resolver by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39260](https://github.com/BerriAI/litellm/pull/39260)
- feat(mcp): semantic tool search for the native MCP Gateway by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39404](https://github.com/BerriAI/litellm/pull/39404)
- fix(logging): redact credential query params from the uvicorn access log by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39293](https://github.com/BerriAI/litellm/pull/39293)
- feat(model\_prices): add meta/muse-spark-1.3 and its contributor tier by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39417](https://github.com/BerriAI/litellm/pull/39417)
- refactor(core): move audio transcription into core by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39126](https://github.com/BerriAI/litellm/pull/39126)
- fix(proxy): build coordination Redis from REDIS\_\* env vars unconditionally by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39410](https://github.com/BerriAI/litellm/pull/39410)
- test: add interactive Rust Python parity harness by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39419](https://github.com/BerriAI/litellm/pull/39419)
- test(proxy): verify NO\_DOCS/NO\_REDOC/NO\_OPENAPI restrict every doc surface by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39378](https://github.com/BerriAI/litellm/pull/39378)
- test(bedrock): accept the router kwarg in the knowledge base search fake by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39420](https://github.com/BerriAI/litellm/pull/39420)
- refactor(python-bridge): split routes and add shared function tracing by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39031](https://github.com/BerriAI/litellm/pull/39031)
- fix(python-bridge): harden sync and async execution boundaries by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39332](https://github.com/BerriAI/litellm/pull/39332)
- refactor(python-bridge): declare sync and async routes once by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39333](https://github.com/BerriAI/litellm/pull/39333)
- feat(python): unify Rust opt-in and bridge policy by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39334](https://github.com/BerriAI/litellm/pull/39334)
- feat(router): add heuristic v2 complexity routing by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39276](https://github.com/BerriAI/litellm/pull/39276)
- fix(anthropic): upgrade legacy thinking to adaptive on adaptive-only Claude models for chat, Bedrock Converse, Invoke, Vertex AI, and Databricks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39159](https://github.com/BerriAI/litellm/pull/39159)
- fix(proxy): mark session/SSO/SAML cookies Secure behind a TLS-terminating reverse proxy by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39391](https://github.com/BerriAI/litellm/pull/39391)
- fix(bedrock): honor BEDROCK\_MANTLE\_API\_BASE on bedrock/mantle messages and chat URLs by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39364](https://github.com/BerriAI/litellm/pull/39364)
- fix(bedrock): strip client\_metadata from converse additionalModelRequestFields by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35967](https://github.com/BerriAI/litellm/pull/35967)
- chore(techdebt): clear fresh debt from the 2026-08-31 and 2026-09-01 windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39091](https://github.com/BerriAI/litellm/pull/39091)
- fix(mcp): cap tools preview and test-connection at the listing timeout and name the unreachable upstream by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38791](https://github.com/BerriAI/litellm/pull/38791)
- fix(hosted\_vllm): forward truncate\_prompt\_tokens on rerank requests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39363](https://github.com/BerriAI/litellm/pull/39363)
- fix(messages): drop cache\_control ttl on non-Anthropic /v1/messages passthrough by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39355](https://github.com/BerriAI/litellm/pull/39355)
- fix(bedrock\_mantle): carry per-request AWS credentials into chat completions SigV4 signing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39362](https://github.com/BerriAI/litellm/pull/39362)
- feat(router): add a hybrid classifier that defers near tier boundaries by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39403](https://github.com/BerriAI/litellm/pull/39403)
- fix: recover the v2 migration resolver from concurrent migrate deploy deadlocks by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39187](https://github.com/BerriAI/litellm/pull/39187)
- fix(ollama\_chat): stamp finish\_reason tool\_calls when tool calls streamed before the done chunk by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39010](https://github.com/BerriAI/litellm/pull/39010)
- fix(router): route Claude Code subagents through session router by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39239](https://github.com/BerriAI/litellm/pull/39239)
- fix(responses): keep namespace tools intact when a guardrail returns them unchanged by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39366](https://github.com/BerriAI/litellm/pull/39366)
- fix(vector-store): resolve embedding credentials per request by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;38936](https://github.com/BerriAI/litellm/pull/38936)
- test(e2e/ui): give the seeded users passwords that pass the default password policy by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39442](https://github.com/BerriAI/litellm/pull/39442)
- fix(http\_handler): honor HTTP(S)\_PROXY / NO\_PROXY when force\_ipv4 uses the httpx transport by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39443](https://github.com/BerriAI/litellm/pull/39443)
- fix(proxy): stop leaking internal exception details to clients by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39380](https://github.com/BerriAI/litellm/pull/39380)
- fix(guardrails): forward mode and streaming params to crowdstrike\_aidr handler by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39317](https://github.com/BerriAI/litellm/pull/39317)
- fix(mcp): gate the connect-time OBO pre-flight on the key's allowed servers by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39447](https://github.com/BerriAI/litellm/pull/39447)
- fix(responses): keep provider response headers in streaming logging callbacks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38131](https://github.com/BerriAI/litellm/pull/38131)
- fix(mcp): fence an outbound-token write against an overlapping invalidation by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35398](https://github.com/BerriAI/litellm/pull/35398)
- feat(cli): pre-fill the SSO verification code in the browser when the proxy allows it by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39428](https://github.com/BerriAI/litellm/pull/39428)
- fix(ui): paginate request logs by session groups server-side by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39257](https://github.com/BerriAI/litellm/pull/39257)
- feat(proxy): serve the auto-router preset catalog at runtime by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39412](https://github.com/BerriAI/litellm/pull/39412)
- docs: define Rust Python harness structure by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39456](https://github.com/BerriAI/litellm/pull/39456)
- fix(guardrails): apply PUT /guardrails/{id} to the serving worker immediately and reject invalid configs with 422 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38877](https://github.com/BerriAI/litellm/pull/38877)
- test(responses): expect the 404 OpenAI now returns for an unknown model by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39457](https://github.com/BerriAI/litellm/pull/39457)
- fix(guardrails): skip streaming guardrail rounds that re-scan cleared output by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39386](https://github.com/BerriAI/litellm/pull/39386)
- fix: keep litellm importable on Python 3.10 and guard 3.11-only typing imports in CI by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39448](https://github.com/BerriAI/litellm/pull/39448)
- fix(proxy): keep SpendLogs and callback session ids in sync when the request has none by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39450](https://github.com/BerriAI/litellm/pull/39450)
- feat(router): arm safeguard-refusal fallback on generic chains when no content-policy list exists by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39274](https://github.com/BerriAI/litellm/pull/39274)
- feat(azure): support credential chain for storage by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39229](https://github.com/BerriAI/litellm/pull/39229)
- chore(crowdstrike): expect the deduped end-of-stream scan in crowdstrike cadence test by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39467](https://github.com/BerriAI/litellm/pull/39467)
- fix(model\_armor): handle Anthropic Messages and Responses streams in post\_call by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39181](https://github.com/BerriAI/litellm/pull/39181)
- test: add OCR python-to-rust test parity ledger (WIP) by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39434](https://github.com/BerriAI/litellm/pull/39434)
- feat(complexity\_router): opt-in modality override of a kept session-affinity pin by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39454](https://github.com/BerriAI/litellm/pull/39454)
- feat(datadog\_llm\_obs): cost tag dimensions, router decision fields, reasoning token metric, redaction gating by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39402](https://github.com/BerriAI/litellm/pull/39402)
- test(rust-python-harness): wire existing e2e SDK tests into the matrix by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39463](https://github.com/BerriAI/litellm/pull/39463)
- fix(mcp): never exchange the LiteLLM virtual key as the upstream subject token by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39446](https://github.com/BerriAI/litellm/pull/39446)
- test: add mistral ocr transformation parity coverage by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39482](https://github.com/BerriAI/litellm/pull/39482)
- test(vector-store): accept embedding\_executor in the Bedrock KB hook fake handler by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39472](https://github.com/BerriAI/litellm/pull/39472)
- refactor(s3\_vectors): embed search queries through the shared vector store executor by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39474](https://github.com/BerriAI/litellm/pull/39474)
- fix(xai): bill from the cost xAI reports instead of recomputing it (internal copy of [#&#8203;36281](https://github.com/BerriAI/litellm/issues/36281)) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39441](https://github.com/BerriAI/litellm/pull/39441)
- feat(ui): add 1M context auto-router preset by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39490](https://github.com/BerriAI/litellm/pull/39490)
- fix(ui): stop the create team form resetting organization and models by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39476](https://github.com/BerriAI/litellm/pull/39476)
- fix(ui): read the preset catalog at runtime in the dashboard tests by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39478](https://github.com/BerriAI/litellm/pull/39478)
- fix(sso): resolve multi-valued role claims to the highest privilege role by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39480](https://github.com/BerriAI/litellm/pull/39480)
- fix(guardrail): hide-secrets playground redaction and guardrail telemetry by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39398](https://github.com/BerriAI/litellm/pull/39398)
- fix(test): drop the duplicate embedding\_executor arg in the Bedrock KB fake handler by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39502](https://github.com/BerriAI/litellm/pull/39502)
- fix(ui): keep Virtual Keys list state in the URL so it survives leaving the page by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39481](https://github.com/BerriAI/litellm/pull/39481)
- fix(proxy): 404 a credential delete that matched nothing, and raise instead of return by [@&#8203;eeshsaxena](https://github.com/eeshsaxena) in [#&#8203;36260](https://github.com/BerriAI/litellm/pull/36260)
- fix(proxy-extras): only spend a migrate-deploy attempt when a pass made no progress by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39506](https://github.com/BerriAI/litellm/pull/39506)
- feat(cli): enable Claude Code gateway model discovery by default in lite claude by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39445](https://github.com/BerriAI/litellm/pull/39445)
- fix(docker): bump nginx runtime to 1.31.5-alpine3.24 and pin digest by [@&#8203;rakeshrepository](https://github.com/rakeshrepository) in [#&#8203;39561](https://github.com/BerriAI/litellm/pull/39561)
- fix: 1.99.0-rc2 UI bug batch (empty org on key create, session pagination, access group rename/delete) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39436](https://github.com/BerriAI/litellm/pull/39436)
- feat(auto-router): support classifier reasoning effort by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39372](https://github.com/BerriAI/litellm/pull/39372)
- fix(ui): replace the key detail URL entry when a virtual key is rotated by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39471](https://github.com/BerriAI/litellm/pull/39471)
- test(timeout): time out against the local fake endpoint instead of api.openai.com by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39583](https://github.com/BerriAI/litellm/pull/39583)
- test(harness): add OCR parity with migration strategy runners by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;38765](https://github.com/BerriAI/litellm/pull/38765)
- fix(databricks): strip thinking\_blocks and reasoning\_content from outbound messages by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39409](https://github.com/BerriAI/litellm/pull/39409)
- test(ocr): record provider fixtures in the migration harness by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39425](https://github.com/BerriAI/litellm/pull/39425)
- feat(ui): keyset-paginate request logs by session trace by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38794](https://github.com/BerriAI/litellm/pull/38794)
- fix(proxy/db): translate libpq sslrootcert and verify-\* into Prisma's strict TLS params by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39563](https://github.com/BerriAI/litellm/pull/39563)
- fix(agents): keep the published agent in public\_agent\_groups by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39554](https://github.com/BerriAI/litellm/pull/39554)
- fix(mcp): scope allow-all servers to virtual keys by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39531](https://github.com/BerriAI/litellm/pull/39531)
- fix(team): generate team IDs for blank input by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39571](https://github.com/BerriAI/litellm/pull/39571)
- fix(bedrock\_mantle): stop dropping the web\_search tool on /v1/responses by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35987](https://github.com/BerriAI/litellm/pull/35987)
- chore: bump litellm-enterprise 0.1.63 -> 0.1.64, litellm-proxy-extras 0.4.92 -> 0.4.93 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39595](https://github.com/BerriAI/litellm/pull/39595)
- fix(images): forward gpt-image supported params like background to OpenAI and Azure by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39525](https://github.com/BerriAI/litellm/pull/39525)
- fix(proxy): return persisted team memberships from /user/new so first CLI login gets the default team by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39545](https://github.com/BerriAI/litellm/pull/39545)
- fix(spend\_tracking): add missing\_session\_id: omit to leave SpendLogs.session\_id null without a client session by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39458](https://github.com/BerriAI/litellm/pull/39458)
- fix: stop a cleared Organization field from failing key creation by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39316](https://github.com/BerriAI/litellm/pull/39316)
- fix(ui): show MCP servers and agents inherited from access groups on team overview by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39215](https://github.com/BerriAI/litellm/pull/39215)
- fix(proxy): expose configured mode for auto-router models by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39619](https://github.com/BerriAI/litellm/pull/39619)
- fix(ui): aggregate session token usage in the logs table by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39598](https://github.com/BerriAI/litellm/pull/39598)
- fix(cost): apply off\_peak\_pricing in the dashscope cost calculator by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39592](https://github.com/BerriAI/litellm/pull/39592)
- test(bedrock): drop EOL cohere.command-r-plus-v1:0 from local\_testing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39608](https://github.com/BerriAI/litellm/pull/39608)
- fix(openai): default stream usage on PrivateLink and regional api.openai.com hosts by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39614](https://github.com/BerriAI/litellm/pull/39614)
- fix(proxy): drop anthropic-beta on the Vertex passthrough count-tokens route by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39597](https://github.com/BerriAI/litellm/pull/39597)
- fix(headroom): resolve CCR retrieval on streaming /v1/responses by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38808](https://github.com/BerriAI/litellm/pull/38808)
- fix(openai): bridge gpt-5.4+ tool calls to /v1/responses on every api.openai.com host by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39587](https://github.com/BerriAI/litellm/pull/39587)
- fix(router): pin JWT-authenticated callers by user id in deployment\_affinity by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39594](https://github.com/BerriAI/litellm/pull/39594)
- fix(cost): bill bedrock\_mantle web search at $12 per 1k queries using Bedrock's reported count by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39610](https://github.com/BerriAI/litellm/pull/39610)
- fix(azure\_ai): don't reclassify Foundry deployments as azure provider by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38975](https://github.com/BerriAI/litellm/pull/38975)
- fix(vector\_stores): only list vector stores the caller was granted by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39612](https://github.com/BerriAI/litellm/pull/39612)
- feat(models): add gpt-6-astra pricing and metadata by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39622](https://github.com/BerriAI/litellm/pull/39622)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39593](https://github.com/BerriAI/litellm/pull/39593)
- feat(router): limit heuristic\_v2 auto-routers to one without the auto\_router license feature by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39468](https://github.com/BerriAI/litellm/pull/39468)
- fix(ui): clear agents when updating team permissions by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39600](https://github.com/BerriAI/litellm/pull/39600)
- fix(auto\_router): bill the routing embedding to the caller's key and team by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39532](https://github.com/BerriAI/litellm/pull/39532)
- test(router): cover get\_configured\_mode so router\_code\_coverage passes by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;39630](https://github.com/BerriAI/litellm/pull/39630)
- fix: treat gpt-6 names as the gpt-5 request family in OpenAI and Azure configs by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39631](https://github.com/BerriAI/litellm/pull/39631)
- fix(prompts): key the in-memory prompt registry by environment by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38440](https://github.com/BerriAI/litellm/pull/38440)
- fix(ui): let the Internal Users search box match user\_id as well as email by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39604](https://github.com/BerriAI/litellm/pull/39604)
- test(responses): bound the background stream cancel e2e so an upstream stall skips fast by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39617](https://github.com/BerriAI/litellm/pull/39617)
- fix(vertex): add the API version to versionless project routes on the Vertex passthrough by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39625](https://github.com/BerriAI/litellm/pull/39625)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39648](https://github.com/BerriAI/litellm/pull/39648)
- fix(spend\_tracking): key /v1/messages spend rows on the msg\_ id the client received by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39511](https://github.com/BerriAI/litellm/pull/39511)
- ci(rust): build and test the ai-gateway server feature by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39493](https://github.com/BerriAI/litellm/pull/39493)
- ci(ui): run the UI build check through the image's ui-builder stage by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39496](https://github.com/BerriAI/litellm/pull/39496)
- fix(proxy): parse numeric multipart fields on /v1/images/edits back into numbers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39510](https://github.com/BerriAI/litellm/pull/39510)
- fix(guardrails): remove the module-global translation mapping that leaked between tests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39543](https://github.com/BerriAI/litellm/pull/39543)
- feat(azure\_ai): add grok-4.6 to the model cost map by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39426](https://github.com/BerriAI/litellm/pull/39426)
- fix: attach vector store search\_results when a guardrail is registered by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38984](https://github.com/BerriAI/litellm/pull/38984)
- fix(proxy): stop putting the literal string "None" in error payloads by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39521](https://github.com/BerriAI/litellm/pull/39521)
- fix(router): keep retry breadcrumbs per request and out of the request snapshot by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39491](https://github.com/BerriAI/litellm/pull/39491)
- fix(vector-stores): survive a failing vector store search in the chat completions hook by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39495](https://github.com/BerriAI/litellm/pull/39495)
- fix(utils): redact credential kwargs from the set\_verbose request line by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39526](https://github.com/BerriAI/litellm/pull/39526)
- fix(bedrock): skip the SigV4 credential chain when a bearer token is configured by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39411](https://github.com/BerriAI/litellm/pull/39411)
- fix(proxy-extras): kill the whole Prisma process group when a command times out by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39466](https://github.com/BerriAI/litellm/pull/39466)
- fix(rag): forward the managed vector store's params to the search call by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39452](https://github.com/BerriAI/litellm/pull/39452)
- fix(utils): redact credentials nested in extra\_body on the verbose optional-params…
hbjydev pushed a commit to hbjydev/phoebe that referenced this pull request Sep 20, 2026
…102.0) (#630)

This PR contains the following updates:

| Package | Update | Change |
|---|---|---|
| [ghcr.io/berriai/litellm](https://images.chainguard.dev/directory/image/wolfi-base/overview) ([source](https://github.com/BerriAI/litellm)) | minor | `v1.100.1` → `v1.102.0` |

---

> ⚠️ **Warning**
>
> Some dependencies could not be looked up. Check the [Dependency Dashboard](issues/141) for more information.

---

### Release Notes

<details>
<summary>BerriAI/litellm (ghcr.io/berriai/litellm)</summary>

### [`v1.102.0`](https://github.com/BerriAI/litellm/compare/v1.101.0...v1.102.0)

[Compare Source](https://github.com/BerriAI/litellm/compare/v1.101.0...v1.102.0)

### [`v1.101.0`](https://github.com/BerriAI/litellm/releases/tag/v1.101.0)

[Compare Source](https://github.com/BerriAI/litellm/compare/v1.100.1...v1.101.0)

##### Verify Docker Image Signature

All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d726492c814b3644b7d376029d0).

**Verify using the pinned commit hash (recommended):**

A commit hash is cryptographically immutable, so this is the strongest way to ensure you are using the original signing key:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/0112e53046018d726492c814b3644b7d376029d0/cosign.pub \
  ghcr.io/berriai/litellm:v1.101.0
```

**Verify using the release tag (convenience):**

Tags are protected in this repository and resolve to the same key. This option is easier to read but relies on tag protection rules:

```bash
cosign verify \
  --key https://raw.githubusercontent.com/BerriAI/litellm/v1.101.0/cosign.pub \
  ghcr.io/berriai/litellm:v1.101.0
```

Expected output:

```
The following checks were performed on each of these signatures:
  - The cosign claims were validated
  - The signatures were verified against the specified public key
```

***

##### What's Changed

- fix(proxy): emit timing headers and overhead for /v1/messages and /v1/responses by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38840](https://github.com/BerriAI/litellm/pull/38840)
- fix(tests): derive the no-cache-read-rate savings baseline from the model map by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38863](https://github.com/BerriAI/litellm/pull/38863)
- chore(typing): clear Any seams across 47 files, ratchet basedpyright ceilings -3,302 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;37778](https://github.com/BerriAI/litellm/pull/37778)
- chore(typing): clear 1.2k basedpyright Any errors across 16 hotspot files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;36722](https://github.com/BerriAI/litellm/pull/36722)
- feat(bedrock): honor streaming buffer/sampling config for unbuffered post\_call scans by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38722](https://github.com/BerriAI/litellm/pull/38722)
- feat(cli): set ENABLE\_TOOL\_SEARCH=true for lite claude by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38942](https://github.com/BerriAI/litellm/pull/38942)
- fix(proxy): deliver budget alerts on webhook-only alerting and accept ALERTING\_WEBHOOK\_URL by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38441](https://github.com/BerriAI/litellm/pull/38441)
- docs(claude.md): require tests to check behavior, not code structure by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;38772](https://github.com/BerriAI/litellm/pull/38772)
- chore(newrelic): cover static default\_team\_settings per-team routing by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38857](https://github.com/BerriAI/litellm/pull/38857)
- fix: update stale source URLs and deprecation dates in model cost map by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38801](https://github.com/BerriAI/litellm/pull/38801)
- feat(ci): close duplicate issues after a 3-day grace period by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;38381](https://github.com/BerriAI/litellm/pull/38381)
- docs(proxy): clarify spend semantics on /v2/user/info and /user/daily/activity by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38883](https://github.com/BerriAI/litellm/pull/38883)
- fix(guardrails): configure Prompt Security file timeout policy by [@&#8203;davida-ps](https://github.com/davida-ps) in [#&#8203;38083](https://github.com/BerriAI/litellm/pull/38083)
- fix(bedrock): stop duplicating Converse config blocks inside inferenceConfig by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38993](https://github.com/BerriAI/litellm/pull/38993)
- fix(guardrails): exclude images from HiddenLayer v1 scans by [@&#8203;Ashton-Sidhu](https://github.com/Ashton-Sidhu) in [#&#8203;29210](https://github.com/BerriAI/litellm/pull/29210)
- feat(spend\_tracking): persist router metadata in spend logs for internal router models by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39001](https://github.com/BerriAI/litellm/pull/39001)
- fix(vertex\_ai): graft default vertex path when api\_base has a version-only path by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38986](https://github.com/BerriAI/litellm/pull/38986)
- fix(proxy): allow unblocking customers via /customer/update by [@&#8203;cat0825](https://github.com/cat0825) in [#&#8203;34696](https://github.com/BerriAI/litellm/pull/34696)
- feat(openai): support workload identity federation (OIDC token exchange) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38995](https://github.com/BerriAI/litellm/pull/38995)
- fix(otel): emit cache token counts on OTel v2 LLM spans by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38716](https://github.com/BerriAI/litellm/pull/38716)
- feat(proxy): add /v1/responses/input\_tokens token counting endpoint by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38997](https://github.com/BerriAI/litellm/pull/38997)
- fix(docker): bump wolfi-base for glibc 2.44 and pin apk python to 3.13 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38917](https://github.com/BerriAI/litellm/pull/38917)
- fix(docker): bump wolfi-base for glibc 2.44 and pin apk python to 3.13 in migrations image by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38973](https://github.com/BerriAI/litellm/pull/38973)
- feat(friendli): add zai-org/GLM-5.3-Flash model pricing by [@&#8203;Lee-Si-Yoon](https://github.com/Lee-Si-Yoon) in [#&#8203;38880](https://github.com/BerriAI/litellm/pull/38880)
- chore(techdebt): clear fresh debt from the 2026-08-29 and 2026-08-30 windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38884](https://github.com/BerriAI/litellm/pull/38884)
- fix(bedrock): surface Nova Sonic user transcripts, speech events, and usage in realtime API by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38597](https://github.com/BerriAI/litellm/pull/38597)
- fix(guardrails): carry Anthropic url image sources through to guardrails by [@&#8203;samtsai15](https://github.com/samtsai15) in [#&#8203;38940](https://github.com/BerriAI/litellm/pull/38940)
- feat(friendli): add zai-org/GLM-5.3 model pricing by [@&#8203;Lee-Si-Yoon](https://github.com/Lee-Si-Yoon) in [#&#8203;38881](https://github.com/BerriAI/litellm/pull/38881)
- fix(router): apply model renames to the in-memory deployment list by [@&#8203;yatishgoel](https://github.com/yatishgoel) in [#&#8203;38479](https://github.com/BerriAI/litellm/pull/38479)
- test(e2e): cover SCIM token creation and SCIM API auth in the Admin UI suite by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39027](https://github.com/BerriAI/litellm/pull/39027)
- feat(gigachat): add native API passthrough routes with spend logging by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38913](https://github.com/BerriAI/litellm/pull/38913)
- feat(gigachat): add passthrough gigachat route by [@&#8203;KnyazSh](https://github.com/KnyazSh) in [#&#8203;25886](https://github.com/BerriAI/litellm/pull/25886)
- feat(complexity-router): add classification\_mode to skip classifier on continuation turns by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38861](https://github.com/BerriAI/litellm/pull/38861)
- fix(proxy): preserve model table columns on master key rotation by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38878](https://github.com/BerriAI/litellm/pull/38878)
- fix(speech): stop forwarding response\_format as a chat param for Gemini TTS by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38819](https://github.com/BerriAI/litellm/pull/38819)
- fix(proxy): return 200 from /model/block and /model/unblock instead of 500 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38873](https://github.com/BerriAI/litellm/pull/38873)
- feat(complexity\_router): escalate oversized prompts to a tier that fits before dispatch by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;38844](https://github.com/BerriAI/litellm/pull/38844)
- feat(shadow\_eval): target teams and users so JWT-auth traffic can be evaluated by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39015](https://github.com/BerriAI/litellm/pull/39015)
- fix(anthropic\_messages): drain upstream in a detached pump so client … by [@&#8203;nuernber](https://github.com/nuernber) in [#&#8203;36008](https://github.com/BerriAI/litellm/pull/36008)
- refactor(proxy): bound the budget window seed by time instead of request ids by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;38851](https://github.com/BerriAI/litellm/pull/38851)
- fix(proxy): ship psycopg so partitioned SpendLogs detection actually runs by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;38994](https://github.com/BerriAI/litellm/pull/38994)
- test(e2e): assert user-observable behavior instead of DOM structure by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39016](https://github.com/BerriAI/litellm/pull/39016)
- build(rust): configure native extension profiles by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39020](https://github.com/BerriAI/litellm/pull/39020)
- fix(ui): keep litellm\_credential\_name from LiteLLM Params JSON when no credential is selected by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39005](https://github.com/BerriAI/litellm/pull/39005)
- Revert "fix(ui): keep litellm\_credential\_name from LiteLLM Params JSON when no credential is selected" by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39046](https://github.com/BerriAI/litellm/pull/39046)
- fix(auth): quiet malformed virtual key rejections to stdout by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38838](https://github.com/BerriAI/litellm/pull/38838)
- fix(proxy): wire team-level logging callbacks into passthrough endpoints by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;38979](https://github.com/BerriAI/litellm/pull/38979)
- feat(complexity\_router): opt-in modality-based capability routing for image requests by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39032](https://github.com/BerriAI/litellm/pull/39032)
- fix(ui): let the auto-router scoring tier list follow the theme by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39040](https://github.com/BerriAI/litellm/pull/39040)
- feat(ui): auto-router controls for context-window escalation by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39054](https://github.com/BerriAI/litellm/pull/39054)
- fix(redis): coerce env var string types and fix param discovery through decorator wrappers by [@&#8203;koladefaj](https://github.com/koladefaj) in [#&#8203;30644](https://github.com/BerriAI/litellm/pull/30644)
- feat(key management): show budget window usage on /key/info by [@&#8203;Thijmen](https://github.com/Thijmen) in [#&#8203;37044](https://github.com/BerriAI/litellm/pull/37044)
- fix(websearch): reject invalid explicit search tool selections by [@&#8203;georgeatparallel](https://github.com/georgeatparallel) in [#&#8203;38113](https://github.com/BerriAI/litellm/pull/38113)
- feat(shadow\_eval): compare several auto-routers on one job's sampled traffic by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39028](https://github.com/BerriAI/litellm/pull/39028)
- fix(speech): honor pcm/wav response\_format for Gemini TTS and reject unsupported containers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38868](https://github.com/BerriAI/litellm/pull/38868)
- fix(proxy): match /v1/audio/speech content-type to the returned audio format by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38798](https://github.com/BerriAI/litellm/pull/38798)
- test(e2e): drop the two mgmt registry cells no shared-proxy test can cover by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39055](https://github.com/BerriAI/litellm/pull/39055)
- feat(ui): one classification frequency picker for complexity auto-routers by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39042](https://github.com/BerriAI/litellm/pull/39042)
- test(e2e/ui): automate 8 manual QA checklist flows by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39025](https://github.com/BerriAI/litellm/pull/39025)
- fix(key\_management): allow non-admin key\_type preset transitions on /key/update by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39051](https://github.com/BerriAI/litellm/pull/39051)
- chore(typing): clear 1.1k basedpyright Any errors across 53 backend files by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38796](https://github.com/BerriAI/litellm/pull/38796)
- fix(openai): forward reasoning\_effort for unknown model aliases instead of failing closed by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39065](https://github.com/BerriAI/litellm/pull/39065)
- test(e2e-ui): poll credential availability before Test Connect to deflake multi-instance runs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39073](https://github.com/BerriAI/litellm/pull/39073)
- feat(ui): modality routing toggle on the auto-router create and edit forms by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39059](https://github.com/BerriAI/litellm/pull/39059)
- fix(proxy): include litellm\_model\_table in GET /v2/team/list by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39045](https://github.com/BerriAI/litellm/pull/39045)
- fix(bedrock): mask signed request headers in guardrail debug log by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39044](https://github.com/BerriAI/litellm/pull/39044)
- fix(bedrock): forward aws\_external\_id in files and batches credential loading by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39066](https://github.com/BerriAI/litellm/pull/39066)
- fix(mcp): persist alias MCP grants verbatim instead of rewriting to local server ids by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39119](https://github.com/BerriAI/litellm/pull/39119)
- fix(responses): json-encode object tool call arguments in the chat completions bridge by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35417](https://github.com/BerriAI/litellm/pull/35417)
- fix(cost): bill OCR annotation pages via annotation\_cost\_per\_page by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38985](https://github.com/BerriAI/litellm/pull/38985)
- fix(policy\_engine): restore request guardrails list after pipeline allow by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39038](https://github.com/BerriAI/litellm/pull/39038)
- fix(embeddings): omit encoding\_format when the client omits it on OpenAI-compatible calls by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38774](https://github.com/BerriAI/litellm/pull/38774)
- test: deflake MCP registry state, savings cost map, and MCP identity env reload tests by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38891](https://github.com/BerriAI/litellm/pull/38891)
- feat(helm): add Argo CD PreSync hook and rollout strategy knobs to the componentized chart by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39112](https://github.com/BerriAI/litellm/pull/39112)
- fix(registry): veo 3.1 pricing tiers + roll up open registry PRs (glm-5.2, Qwen3.8-Flash, gemma-4-31b, scribe\_v2, fireworks/databricks deepseek v4) + deprecation dates by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38990](https://github.com/BerriAI/litellm/pull/38990)
- test(ui): budget DOM-structure assertions in dashboard tests by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39082](https://github.com/BerriAI/litellm/pull/39082)
- test(ui): assert DataTable behavior instead of DOM structure by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39084](https://github.com/BerriAI/litellm/pull/39084)
- test(ui): query the screen instead of the render result by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39085](https://github.com/BerriAI/litellm/pull/39085)
- fix(ui): stop checkboxes stretching to the full width of a form field by [@&#8203;yatishgoel](https://github.com/yatishgoel) in [#&#8203;39108](https://github.com/BerriAI/litellm/pull/39108)
- chore: bump litellm-enterprise 0.1.62 -> 0.1.63, litellm-proxy-extras 0.4.91 -> 0.4.92, litellm 1.100.0 -> 1.101.0 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39140](https://github.com/BerriAI/litellm/pull/39140)
- revert: restore search tool fallback when no router is configured by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39146](https://github.com/BerriAI/litellm/pull/39146)
- test(websearch): register configured search tool in pre-request hook test by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39074](https://github.com/BerriAI/litellm/pull/39074)
- feat(proxy): default to the v2 migration resolver, keep v1 as an opt-out by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;31125](https://github.com/BerriAI/litellm/pull/31125)
- build(deps): bump browserslist to 4.28.8 to clear osv-scan by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39142](https://github.com/BerriAI/litellm/pull/39142)
- fix(ui): render the skill detail page with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39130](https://github.com/BerriAI/litellm/pull/39130)
- feat: add Azure AI DeepSeek V4 Flash 0731 pricing by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39023](https://github.com/BerriAI/litellm/pull/39023)
- fix(streaming): keep response id stable across streamed chunks by [@&#8203;Timik232](https://github.com/Timik232) in [#&#8203;38106](https://github.com/BerriAI/litellm/pull/38106)
- test(e2e/ui): cover the Budgets page create, edit and delete flows by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39052](https://github.com/BerriAI/litellm/pull/39052)
- feat(dashscope): add QwenCloud and Qwen AI Platform provider aliases by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39149](https://github.com/BerriAI/litellm/pull/39149)
- fix(bedrock): forward native structured outputs on Invoke instead of silently inlining the schema by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39070](https://github.com/BerriAI/litellm/pull/39070)
- test(e2e/ui): cover creating, testing and deleting a guardrail by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39053](https://github.com/BerriAI/litellm/pull/39053)
- refactor(types): replace Any with precise types across 73 modules by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39104](https://github.com/BerriAI/litellm/pull/39104)
- feat(models): add Claude Fable 5.1 across Anthropic, Bedrock, Vertex AI, and Azure AI by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39148](https://github.com/BerriAI/litellm/pull/39148)
- feat(guardrails): add Alice guardrail by [@&#8203;seanyasno-af](https://github.com/seanyasno-af) in [#&#8203;38898](https://github.com/BerriAI/litellm/pull/38898)
- test(e2e/ui): cover the Logs page filter drawer by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39056](https://github.com/BerriAI/litellm/pull/39056)
- test(e2e/ui): stop the suite failing on things that are not regressions by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39063](https://github.com/BerriAI/litellm/pull/39063)
- test(e2e/ui): cover the team Settings tab by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39058](https://github.com/BerriAI/litellm/pull/39058)
- test(e2e/ui): cover the Usage page activity tabs by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39061](https://github.com/BerriAI/litellm/pull/39061)
- fix(ui): render the guardrail garden detail page with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39131](https://github.com/BerriAI/litellm/pull/39131)
- fix(responses): tool call id shape breaks gpt-5 -> claude fallback conversations by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39144](https://github.com/BerriAI/litellm/pull/39144)
- fix(openai): drop tool\_choice when request has no tools on chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39147](https://github.com/BerriAI/litellm/pull/39147)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39141](https://github.com/BerriAI/litellm/pull/39141)
- test(ui): pick select options by role instead of by text by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39175](https://github.com/BerriAI/litellm/pull/39175)
- feat(cost): support time-based off-peak pricing in cost calculation by [@&#8203;Srivatsa03](https://github.com/Srivatsa03) in [#&#8203;31725](https://github.com/BerriAI/litellm/pull/31725)
- fix(openai): flatten top-level tool schema combinators on chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38839](https://github.com/BerriAI/litellm/pull/38839)
- fix(s3): bound s3 object keys and download filenames for long Responses API ids by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39164](https://github.com/BerriAI/litellm/pull/39164)
- revert: default the proxy back to the v1 migration resolver by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39178](https://github.com/BerriAI/litellm/pull/39178)
- fix(prometheus): bound requested\_model label cardinality on client failure paths by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39136](https://github.com/BerriAI/litellm/pull/39136)
- feat(ui): add search to the Agent Hub tab and admin agents table by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39155](https://github.com/BerriAI/litellm/pull/39155)
- fix(anthropic): fix response\_format for claude-fable-5-1 on Vertex AI and Bedrock by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39184](https://github.com/BerriAI/litellm/pull/39184)
- fix: keep litellm\_credential\_name from LiteLLM Params JSON and gate stored credential attach to proxy admins by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39047](https://github.com/BerriAI/litellm/pull/39047)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39186](https://github.com/BerriAI/litellm/pull/39186)
- test: exempt MockTransport request-shape embedding tests from VCR replay by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39185](https://github.com/BerriAI/litellm/pull/39185)
- fix(ui): render the logs Tools panel with theme tokens by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39129](https://github.com/BerriAI/litellm/pull/39129)
- fix(proxy): default max\_idle\_connection\_lifetime to 60s on DB URLs by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39134](https://github.com/BerriAI/litellm/pull/39134)
- fix(mcp): follow tools/list pagination from upstream servers by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39172](https://github.com/BerriAI/litellm/pull/39172)
- fix(proxy): resolve router model aliases in /utils/supported\_openai\_params by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39000](https://github.com/BerriAI/litellm/pull/39000)
- fix(azure): flatten top-level tool schema combinators on Azure chat completions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38870](https://github.com/BerriAI/litellm/pull/38870)
- fix(bedrock): route streamed responses-API output through the unified guardrail by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38734](https://github.com/BerriAI/litellm/pull/38734)
- fix(ui): hide model write affordances from view-only admin sessions by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38872](https://github.com/BerriAI/litellm/pull/38872)
- fix(cli): quote the Claude Code apiKeyHelper for cmd.exe on Windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39174](https://github.com/BerriAI/litellm/pull/39174)
- fix(logging): guarantee max\_parallel\_requests slot release when streaming logging fails by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39093](https://github.com/BerriAI/litellm/pull/39093)
- feat(alerting): slack alerts for per-user daily/monthly spend thresholds and spend anomaly detection by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38438](https://github.com/BerriAI/litellm/pull/38438)
- fix(docker): add public Wolfi apk repo to runtime image by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39033](https://github.com/BerriAI/litellm/pull/39033)
- fix(router): keep order fallback on the requested order level by [@&#8203;emerzon](https://github.com/emerzon) in [#&#8203;38969](https://github.com/BerriAI/litellm/pull/38969)
- test(e2e): cover retry-on-timeout and the context-window fallback by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39197](https://github.com/BerriAI/litellm/pull/39197)
- fix(budget): reject known estimates over remaining budget under fail\_closed\_budget\_enforcement by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39214](https://github.com/BerriAI/litellm/pull/39214)
- fix: stop a cleared Team field from blocking personal key creation by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39206](https://github.com/BerriAI/litellm/pull/39206)
- test: record each e2e test's source location in the JUnit report by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39209](https://github.com/BerriAI/litellm/pull/39209)
- feat(router): fall back on anthropic safeguard refusals on /v1/messages by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39157](https://github.com/BerriAI/litellm/pull/39157)
- fix(proxy): report requested model on Anthropic streaming message\_start by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35816](https://github.com/BerriAI/litellm/pull/35816)
- fix(helm): reuse the generated master key Secret on helm upgrade by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39219](https://github.com/BerriAI/litellm/pull/39219)
- fix(mcp): report per-server outcomes in aggregate REST tools/list by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39232](https://github.com/BerriAI/litellm/pull/39232)
- fix(cost-map): retry transient boot fetch failures and recover config deployments dropped by a stale cost map by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39230](https://github.com/BerriAI/litellm/pull/39230)
- perf(scim): resolve group members with one user table read per member by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39228](https://github.com/BerriAI/litellm/pull/39228)
- fix(docker): install bedrock-realtime extra in monolith proxy images by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39223](https://github.com/BerriAI/litellm/pull/39223)
- fix(aiohttp\_transport): map transport-internal CancelledError to a retryable ConnectError by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39240](https://github.com/BerriAI/litellm/pull/39240)
- fix(bedrock): gate Converse cachePoint emission on model prompt caching support by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39210](https://github.com/BerriAI/litellm/pull/39210)
- fix(datadog\_llm\_obs): send tool calls, tool results and cache tokens in DD's own fields by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39222](https://github.com/BerriAI/litellm/pull/39222)
- feat(prometheus): expose per-key and per-team rate limit allowed and used gauges by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39236](https://github.com/BerriAI/litellm/pull/39236)
- feat(scim): add placeholder listing and merge so a shadowed account can be healed by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39231](https://github.com/BerriAI/litellm/pull/39231)
- fix: normalize provider-specific cache token fields in OTel v2 usage by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39202](https://github.com/BerriAI/litellm/pull/39202)
- fix: stop deployment default API key limits leaking into provider requests by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39211](https://github.com/BerriAI/litellm/pull/39211)
- fix(proxy): keep passthrough logging metadata and model\_info dicts when team callbacks are wired by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39216](https://github.com/BerriAI/litellm/pull/39216)
- fix(guardrails): deliver modify\_response block as valid SSE on streaming chat and Responses by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39036](https://github.com/BerriAI/litellm/pull/39036)
- fix(search): forward search-tool params through the router, complete Parallel AI v1 param mapping by [@&#8203;jliounis](https://github.com/jliounis) in [#&#8203;37883](https://github.com/BerriAI/litellm/pull/37883)
- fix(bedrock): stop Converse crashing on bearer-token auth without SigV4 credentials by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39166](https://github.com/BerriAI/litellm/pull/39166)
- fix(docker): install saml extra in litellm-backend image by [@&#8203;ojensen-berri](https://github.com/ojensen-berri) in [#&#8203;39291](https://github.com/BerriAI/litellm/pull/39291)
- fix(guardrails): run apply\_guardrail-only providers in logging\_only mode by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39297](https://github.com/BerriAI/litellm/pull/39297)
- feat(gemini): day-0 pricing for gemini-3.8-flash by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39340](https://github.com/BerriAI/litellm/pull/39340)
- fix(vertex): avoid duplicate DeepSeek OCR model namespace by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39194](https://github.com/BerriAI/litellm/pull/39194)
- feat(streaming): carry final response cost on streamed usage by default by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39069](https://github.com/BerriAI/litellm/pull/39069)
- fix(rerank): map provider errors with the resolved provider on sync and async paths by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39176](https://github.com/BerriAI/litellm/pull/39176)
- test(e2e): read JUnit properties off the real collected pytest Item by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39246](https://github.com/BerriAI/litellm/pull/39246)
- feat(proxy): configurable display\_name for the Anthropic-shaped /v1/models listing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39238](https://github.com/BerriAI/litellm/pull/39238)
- fix(helm): scale the classic chart's HPA out at the documented 60 percent CPU by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35975](https://github.com/BerriAI/litellm/pull/35975)
- fix(gemini): return enabled thinking content by default by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39160](https://github.com/BerriAI/litellm/pull/39160)
- fix: run access group key sync UPDATEs on the writer, not the read replica by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39128](https://github.com/BerriAI/litellm/pull/39128)
- fix(models): registry audit 2026-09-01: openai realtime and long-context tiers, mistral aliases, voyage, xai, fireworks, together, scaleway, azure ai, govcloud, azure gov, cloudflare whisper, deprecation dates by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39170](https://github.com/BerriAI/litellm/pull/39170)
- fix: apply optional\_pre\_call\_checks and reject unsupported router settings on /config/update by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39249](https://github.com/BerriAI/litellm/pull/39249)
- fix(vector\_stores): s3 vectors search router bypass + rag query config drop + ui error swallow by [@&#8203;michelligabriele](https://github.com/michelligabriele) in [#&#8203;34788](https://github.com/BerriAI/litellm/pull/34788)
- fix(models): key Azure DeepSeek V4 Flash 0731 by its Foundry catalog id by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39341](https://github.com/BerriAI/litellm/pull/39341)
- fix(deps): raise the tornado and pypdf floors for six new advisories by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39188](https://github.com/BerriAI/litellm/pull/39188)
- fix(headroom): stop re-compressing retrieved CCR content in client tool loops by [@&#8203;QuantumBreakz](https://github.com/QuantumBreakz) in [#&#8203;38591](https://github.com/BerriAI/litellm/pull/38591)
- feat(agentcore-a2a): derive runtime session id from A2A message.contextId by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39371](https://github.com/BerriAI/litellm/pull/39371)
- fix(proxy): share per-model budget counters across replicas through the spend counter cache by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39375](https://github.com/BerriAI/litellm/pull/39375)
- fix(proxy-extras): give prisma migrate deploy its own timeout budget by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39365](https://github.com/BerriAI/litellm/pull/39365)
- fix(proxy): route container create and list through model\_list deployments by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39220](https://github.com/BerriAI/litellm/pull/39220)
- test(build): validate release wheel contracts by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39021](https://github.com/BerriAI/litellm/pull/39021)
- refactor(rust): extract domain-neutral Python interop by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39026](https://github.com/BerriAI/litellm/pull/39026)
- refactor(rust): standardize the core Error type by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39331](https://github.com/BerriAI/litellm/pull/39331)
- fix(ui): preserve full AgentCore runtime ARN in agent edit form by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39382](https://github.com/BerriAI/litellm/pull/39382)
- feat(ui): update OpenAI preset model tiers by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39396](https://github.com/BerriAI/litellm/pull/39396)
- fix(router): resolve realtime session model to routed deployment by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;36811](https://github.com/BerriAI/litellm/pull/36811)
- fix(security): restrict and validate file uploads at /v1/files and /upload/logo by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39379](https://github.com/BerriAI/litellm/pull/39379)
- feat(auth): enforce configurable password policy and SSO-only login by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39381](https://github.com/BerriAI/litellm/pull/39381)
- fix(agents): redact secret litellm\_params fields from all /v1/agents responses by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39389](https://github.com/BerriAI/litellm/pull/39389)
- fix(otel): stamp Langfuse root observation input and output from the request task by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39369](https://github.com/BerriAI/litellm/pull/39369)
- fix(guardrails): track and tear down presidio sibling callbacks on delete and update by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39271](https://github.com/BerriAI/litellm/pull/39271)
- fix(spend): keep every-deployment scope on gateway cache-injection marks by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39241](https://github.com/BerriAI/litellm/pull/39241)
- fix(proxy/db): keep prisma predicates from raising TypeError under a mocked prisma module by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39253](https://github.com/BerriAI/litellm/pull/39253)
- fix(proxy): word database 503s by whether the fault is transient by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39256](https://github.com/BerriAI/litellm/pull/39256)
- refactor(utils): remove the dead get\_api\_key provider-key resolver by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39260](https://github.com/BerriAI/litellm/pull/39260)
- feat(mcp): semantic tool search for the native MCP Gateway by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39404](https://github.com/BerriAI/litellm/pull/39404)
- fix(logging): redact credential query params from the uvicorn access log by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39293](https://github.com/BerriAI/litellm/pull/39293)
- feat(model\_prices): add meta/muse-spark-1.3 and its contributor tier by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39417](https://github.com/BerriAI/litellm/pull/39417)
- refactor(core): move audio transcription into core by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39126](https://github.com/BerriAI/litellm/pull/39126)
- fix(proxy): build coordination Redis from REDIS\_\* env vars unconditionally by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39410](https://github.com/BerriAI/litellm/pull/39410)
- test: add interactive Rust Python parity harness by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39419](https://github.com/BerriAI/litellm/pull/39419)
- test(proxy): verify NO\_DOCS/NO\_REDOC/NO\_OPENAPI restrict every doc surface by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39378](https://github.com/BerriAI/litellm/pull/39378)
- test(bedrock): accept the router kwarg in the knowledge base search fake by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39420](https://github.com/BerriAI/litellm/pull/39420)
- refactor(python-bridge): split routes and add shared function tracing by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39031](https://github.com/BerriAI/litellm/pull/39031)
- fix(python-bridge): harden sync and async execution boundaries by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39332](https://github.com/BerriAI/litellm/pull/39332)
- refactor(python-bridge): declare sync and async routes once by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39333](https://github.com/BerriAI/litellm/pull/39333)
- feat(python): unify Rust opt-in and bridge policy by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39334](https://github.com/BerriAI/litellm/pull/39334)
- feat(router): add heuristic v2 complexity routing by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39276](https://github.com/BerriAI/litellm/pull/39276)
- fix(anthropic): upgrade legacy thinking to adaptive on adaptive-only Claude models for chat, Bedrock Converse, Invoke, Vertex AI, and Databricks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39159](https://github.com/BerriAI/litellm/pull/39159)
- fix(proxy): mark session/SSO/SAML cookies Secure behind a TLS-terminating reverse proxy by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39391](https://github.com/BerriAI/litellm/pull/39391)
- fix(bedrock): honor BEDROCK\_MANTLE\_API\_BASE on bedrock/mantle messages and chat URLs by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39364](https://github.com/BerriAI/litellm/pull/39364)
- fix(bedrock): strip client\_metadata from converse additionalModelRequestFields by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35967](https://github.com/BerriAI/litellm/pull/35967)
- chore(techdebt): clear fresh debt from the 2026-08-31 and 2026-09-01 windows by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39091](https://github.com/BerriAI/litellm/pull/39091)
- fix(mcp): cap tools preview and test-connection at the listing timeout and name the unreachable upstream by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38791](https://github.com/BerriAI/litellm/pull/38791)
- fix(hosted\_vllm): forward truncate\_prompt\_tokens on rerank requests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39363](https://github.com/BerriAI/litellm/pull/39363)
- fix(messages): drop cache\_control ttl on non-Anthropic /v1/messages passthrough by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39355](https://github.com/BerriAI/litellm/pull/39355)
- fix(bedrock\_mantle): carry per-request AWS credentials into chat completions SigV4 signing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39362](https://github.com/BerriAI/litellm/pull/39362)
- feat(router): add a hybrid classifier that defers near tier boundaries by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39403](https://github.com/BerriAI/litellm/pull/39403)
- fix: recover the v2 migration resolver from concurrent migrate deploy deadlocks by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39187](https://github.com/BerriAI/litellm/pull/39187)
- fix(ollama\_chat): stamp finish\_reason tool\_calls when tool calls streamed before the done chunk by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39010](https://github.com/BerriAI/litellm/pull/39010)
- fix(router): route Claude Code subagents through session router by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39239](https://github.com/BerriAI/litellm/pull/39239)
- fix(responses): keep namespace tools intact when a guardrail returns them unchanged by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39366](https://github.com/BerriAI/litellm/pull/39366)
- fix(vector-store): resolve embedding credentials per request by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;38936](https://github.com/BerriAI/litellm/pull/38936)
- test(e2e/ui): give the seeded users passwords that pass the default password policy by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39442](https://github.com/BerriAI/litellm/pull/39442)
- fix(http\_handler): honor HTTP(S)\_PROXY / NO\_PROXY when force\_ipv4 uses the httpx transport by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39443](https://github.com/BerriAI/litellm/pull/39443)
- fix(proxy): stop leaking internal exception details to clients by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;39380](https://github.com/BerriAI/litellm/pull/39380)
- fix(guardrails): forward mode and streaming params to crowdstrike\_aidr handler by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39317](https://github.com/BerriAI/litellm/pull/39317)
- fix(mcp): gate the connect-time OBO pre-flight on the key's allowed servers by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39447](https://github.com/BerriAI/litellm/pull/39447)
- fix(responses): keep provider response headers in streaming logging callbacks by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38131](https://github.com/BerriAI/litellm/pull/38131)
- fix(mcp): fence an outbound-token write against an overlapping invalidation by [@&#8203;yassin-berriai](https://github.com/yassin-berriai) in [#&#8203;35398](https://github.com/BerriAI/litellm/pull/35398)
- feat(cli): pre-fill the SSO verification code in the browser when the proxy allows it by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39428](https://github.com/BerriAI/litellm/pull/39428)
- fix(ui): paginate request logs by session groups server-side by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39257](https://github.com/BerriAI/litellm/pull/39257)
- feat(proxy): serve the auto-router preset catalog at runtime by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39412](https://github.com/BerriAI/litellm/pull/39412)
- docs: define Rust Python harness structure by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39456](https://github.com/BerriAI/litellm/pull/39456)
- fix(guardrails): apply PUT /guardrails/{id} to the serving worker immediately and reject invalid configs with 422 by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38877](https://github.com/BerriAI/litellm/pull/38877)
- test(responses): expect the 404 OpenAI now returns for an unknown model by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39457](https://github.com/BerriAI/litellm/pull/39457)
- fix(guardrails): skip streaming guardrail rounds that re-scan cleared output by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39386](https://github.com/BerriAI/litellm/pull/39386)
- fix: keep litellm importable on Python 3.10 and guard 3.11-only typing imports in CI by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39448](https://github.com/BerriAI/litellm/pull/39448)
- fix(proxy): keep SpendLogs and callback session ids in sync when the request has none by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39450](https://github.com/BerriAI/litellm/pull/39450)
- feat(router): arm safeguard-refusal fallback on generic chains when no content-policy list exists by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39274](https://github.com/BerriAI/litellm/pull/39274)
- feat(azure): support credential chain for storage by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39229](https://github.com/BerriAI/litellm/pull/39229)
- chore(crowdstrike): expect the deduped end-of-stream scan in crowdstrike cadence test by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39467](https://github.com/BerriAI/litellm/pull/39467)
- fix(model\_armor): handle Anthropic Messages and Responses streams in post\_call by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39181](https://github.com/BerriAI/litellm/pull/39181)
- test: add OCR python-to-rust test parity ledger (WIP) by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39434](https://github.com/BerriAI/litellm/pull/39434)
- feat(complexity\_router): opt-in modality override of a kept session-affinity pin by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39454](https://github.com/BerriAI/litellm/pull/39454)
- feat(datadog\_llm\_obs): cost tag dimensions, router decision fields, reasoning token metric, redaction gating by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39402](https://github.com/BerriAI/litellm/pull/39402)
- test(rust-python-harness): wire existing e2e SDK tests into the matrix by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39463](https://github.com/BerriAI/litellm/pull/39463)
- fix(mcp): never exchange the LiteLLM virtual key as the upstream subject token by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39446](https://github.com/BerriAI/litellm/pull/39446)
- test: add mistral ocr transformation parity coverage by [@&#8203;ishaan-berri](https://github.com/ishaan-berri) in [#&#8203;39482](https://github.com/BerriAI/litellm/pull/39482)
- test(vector-store): accept embedding\_executor in the Bedrock KB hook fake handler by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39472](https://github.com/BerriAI/litellm/pull/39472)
- refactor(s3\_vectors): embed search queries through the shared vector store executor by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39474](https://github.com/BerriAI/litellm/pull/39474)
- fix(xai): bill from the cost xAI reports instead of recomputing it (internal copy of [#&#8203;36281](https://github.com/BerriAI/litellm/issues/36281)) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39441](https://github.com/BerriAI/litellm/pull/39441)
- feat(ui): add 1M context auto-router preset by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39490](https://github.com/BerriAI/litellm/pull/39490)
- fix(ui): stop the create team form resetting organization and models by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39476](https://github.com/BerriAI/litellm/pull/39476)
- fix(ui): read the preset catalog at runtime in the dashboard tests by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39478](https://github.com/BerriAI/litellm/pull/39478)
- fix(sso): resolve multi-valued role claims to the highest privilege role by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39480](https://github.com/BerriAI/litellm/pull/39480)
- fix(guardrail): hide-secrets playground redaction and guardrail telemetry by [@&#8203;yucheng-berri](https://github.com/yucheng-berri) in [#&#8203;39398](https://github.com/BerriAI/litellm/pull/39398)
- fix(test): drop the duplicate embedding\_executor arg in the Bedrock KB fake handler by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39502](https://github.com/BerriAI/litellm/pull/39502)
- fix(ui): keep Virtual Keys list state in the URL so it survives leaving the page by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39481](https://github.com/BerriAI/litellm/pull/39481)
- fix(proxy): 404 a credential delete that matched nothing, and raise instead of return by [@&#8203;eeshsaxena](https://github.com/eeshsaxena) in [#&#8203;36260](https://github.com/BerriAI/litellm/pull/36260)
- fix(proxy-extras): only spend a migrate-deploy attempt when a pass made no progress by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39506](https://github.com/BerriAI/litellm/pull/39506)
- feat(cli): enable Claude Code gateway model discovery by default in lite claude by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39445](https://github.com/BerriAI/litellm/pull/39445)
- fix(docker): bump nginx runtime to 1.31.5-alpine3.24 and pin digest by [@&#8203;rakeshrepository](https://github.com/rakeshrepository) in [#&#8203;39561](https://github.com/BerriAI/litellm/pull/39561)
- fix: 1.99.0-rc2 UI bug batch (empty org on key create, session pagination, access group rename/delete) by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39436](https://github.com/BerriAI/litellm/pull/39436)
- feat(auto-router): support classifier reasoning effort by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39372](https://github.com/BerriAI/litellm/pull/39372)
- fix(ui): replace the key detail URL entry when a virtual key is rotated by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39471](https://github.com/BerriAI/litellm/pull/39471)
- test(timeout): time out against the local fake endpoint instead of api.openai.com by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39583](https://github.com/BerriAI/litellm/pull/39583)
- test(harness): add OCR parity with migration strategy runners by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;38765](https://github.com/BerriAI/litellm/pull/38765)
- fix(databricks): strip thinking\_blocks and reasoning\_content from outbound messages by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39409](https://github.com/BerriAI/litellm/pull/39409)
- test(ocr): record provider fixtures in the migration harness by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39425](https://github.com/BerriAI/litellm/pull/39425)
- feat(ui): keyset-paginate request logs by session trace by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;38794](https://github.com/BerriAI/litellm/pull/38794)
- fix(proxy/db): translate libpq sslrootcert and verify-\* into Prisma's strict TLS params by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39563](https://github.com/BerriAI/litellm/pull/39563)
- fix(agents): keep the published agent in public\_agent\_groups by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39554](https://github.com/BerriAI/litellm/pull/39554)
- fix(mcp): scope allow-all servers to virtual keys by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39531](https://github.com/BerriAI/litellm/pull/39531)
- fix(team): generate team IDs for blank input by [@&#8203;yujonglee-berri](https://github.com/yujonglee-berri) in [#&#8203;39571](https://github.com/BerriAI/litellm/pull/39571)
- fix(bedrock\_mantle): stop dropping the web\_search tool on /v1/responses by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;35987](https://github.com/BerriAI/litellm/pull/35987)
- chore: bump litellm-enterprise 0.1.63 -> 0.1.64, litellm-proxy-extras 0.4.92 -> 0.4.93 by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39595](https://github.com/BerriAI/litellm/pull/39595)
- fix(images): forward gpt-image supported params like background to OpenAI and Azure by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39525](https://github.com/BerriAI/litellm/pull/39525)
- fix(proxy): return persisted team memberships from /user/new so first CLI login gets the default team by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39545](https://github.com/BerriAI/litellm/pull/39545)
- fix(spend\_tracking): add missing\_session\_id: omit to leave SpendLogs.session\_id null without a client session by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39458](https://github.com/BerriAI/litellm/pull/39458)
- fix: stop a cleared Organization field from failing key creation by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39316](https://github.com/BerriAI/litellm/pull/39316)
- fix(ui): show MCP servers and agents inherited from access groups on team overview by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39215](https://github.com/BerriAI/litellm/pull/39215)
- fix(proxy): expose configured mode for auto-router models by [@&#8203;moe-berri](https://github.com/moe-berri) in [#&#8203;39619](https://github.com/BerriAI/litellm/pull/39619)
- fix(ui): aggregate session token usage in the logs table by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39598](https://github.com/BerriAI/litellm/pull/39598)
- fix(cost): apply off\_peak\_pricing in the dashscope cost calculator by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39592](https://github.com/BerriAI/litellm/pull/39592)
- test(bedrock): drop EOL cohere.command-r-plus-v1:0 from local\_testing by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39608](https://github.com/BerriAI/litellm/pull/39608)
- fix(openai): default stream usage on PrivateLink and regional api.openai.com hosts by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39614](https://github.com/BerriAI/litellm/pull/39614)
- fix(proxy): drop anthropic-beta on the Vertex passthrough count-tokens route by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39597](https://github.com/BerriAI/litellm/pull/39597)
- fix(headroom): resolve CCR retrieval on streaming /v1/responses by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38808](https://github.com/BerriAI/litellm/pull/38808)
- fix(openai): bridge gpt-5.4+ tool calls to /v1/responses on every api.openai.com host by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39587](https://github.com/BerriAI/litellm/pull/39587)
- fix(router): pin JWT-authenticated callers by user id in deployment\_affinity by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39594](https://github.com/BerriAI/litellm/pull/39594)
- fix(cost): bill bedrock\_mantle web search at $12 per 1k queries using Bedrock's reported count by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39610](https://github.com/BerriAI/litellm/pull/39610)
- fix(azure\_ai): don't reclassify Foundry deployments as azure provider by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38975](https://github.com/BerriAI/litellm/pull/38975)
- fix(vector\_stores): only list vector stores the caller was granted by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39612](https://github.com/BerriAI/litellm/pull/39612)
- feat(models): add gpt-6-astra pricing and metadata by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39622](https://github.com/BerriAI/litellm/pull/39622)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39593](https://github.com/BerriAI/litellm/pull/39593)
- feat(router): limit heuristic\_v2 auto-routers to one without the auto\_router license feature by [@&#8203;tin-berri](https://github.com/tin-berri) in [#&#8203;39468](https://github.com/BerriAI/litellm/pull/39468)
- fix(ui): clear agents when updating team permissions by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39600](https://github.com/BerriAI/litellm/pull/39600)
- fix(auto\_router): bill the routing embedding to the caller's key and team by [@&#8203;devin-ai-integration](https://github.com/devin-ai-integration)\[bot] in [#&#8203;39532](https://github.com/BerriAI/litellm/pull/39532)
- test(router): cover get\_configured\_mode so router\_code\_coverage passes by [@&#8203;mubashir1osmani](https://github.com/mubashir1osmani) in [#&#8203;39630](https://github.com/BerriAI/litellm/pull/39630)
- fix: treat gpt-6 names as the gpt-5 request family in OpenAI and Azure configs by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39631](https://github.com/BerriAI/litellm/pull/39631)
- fix(prompts): key the in-memory prompt registry by environment by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38440](https://github.com/BerriAI/litellm/pull/38440)
- fix(ui): let the Internal Users search box match user\_id as well as email by [@&#8203;ryan-crabbe-berri](https://github.com/ryan-crabbe-berri) in [#&#8203;39604](https://github.com/BerriAI/litellm/pull/39604)
- test(responses): bound the background stream cancel e2e so an upstream stall skips fast by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39617](https://github.com/BerriAI/litellm/pull/39617)
- fix(vertex): add the API version to versionless project routes on the Vertex passthrough by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39625](https://github.com/BerriAI/litellm/pull/39625)
- chore(ci): promote internal staging to main by [@&#8203;yuneng-berri](https://github.com/yuneng-berri) in [#&#8203;39648](https://github.com/BerriAI/litellm/pull/39648)
- fix(spend\_tracking): key /v1/messages spend rows on the msg\_ id the client received by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39511](https://github.com/BerriAI/litellm/pull/39511)
- ci(rust): build and test the ai-gateway server feature by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39493](https://github.com/BerriAI/litellm/pull/39493)
- ci(ui): run the UI build check through the image's ui-builder stage by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39496](https://github.com/BerriAI/litellm/pull/39496)
- fix(proxy): parse numeric multipart fields on /v1/images/edits back into numbers by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39510](https://github.com/BerriAI/litellm/pull/39510)
- fix(guardrails): remove the module-global translation mapping that leaked between tests by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39543](https://github.com/BerriAI/litellm/pull/39543)
- feat(azure\_ai): add grok-4.6 to the model cost map by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39426](https://github.com/BerriAI/litellm/pull/39426)
- fix: attach vector store search\_results when a guardrail is registered by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;38984](https://github.com/BerriAI/litellm/pull/38984)
- fix(proxy): stop putting the literal string "None" in error payloads by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39521](https://github.com/BerriAI/litellm/pull/39521)
- fix(router): keep retry breadcrumbs per request and out of the request snapshot by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39491](https://github.com/BerriAI/litellm/pull/39491)
- fix(vector-stores): survive a failing vector store search in the chat completions hook by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39495](https://github.com/BerriAI/litellm/pull/39495)
- fix(utils): redact credential kwargs from the set\_verbose request line by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39526](https://github.com/BerriAI/litellm/pull/39526)
- fix(bedrock): skip the SigV4 credential chain when a bearer token is configured by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39411](https://github.com/BerriAI/litellm/pull/39411)
- fix(proxy-extras): kill the whole Prisma process group when a command times out by [@&#8203;mateo-berri](https://github.com/mateo-berri) in [#&#8203;39466](https://github.com/BerriAI/litellm/pull/39466)
- fix(rag): forward the managed vector store's params to the search call by [@&#8203;mateo-berri](https://github.c…
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants