Repository navigation
Conversation
…ion/mercury-2.5 The inception/mercury-2.5 entry has input and output pricing but is missing cache_read_input_token_cost and supports_prompt_caching, so cached-token usage is not cost-tracked. Inception lists cached input at $0.02 / 1M tokens (standard), and the sibling inception/mercury-2 already carries these fields. Adds: - cache_read_input_token_cost: 2e-08 ($0.02 / 1M standard rate) - supports_prompt_caching: true Applied to both the root cost map and the litellm/ backup copy so they stay in sync. Addresses the cached-pricing portion of BerriAI#40746. Source: https://www.inceptionlabs.ai/models#pricing
Greptile SummaryAdds cached-input pricing and prompt-caching metadata for Inception Mercury 2.5 in both synchronized cost maps
Confidence Score: 4/5This PR is not safe to merge until the active cached-input price is represented and the repository-required regression coverage is added Cached tokens are multiplied directly by the configured rate, so using $0.02/M during the acknowledged $0.004/M promotion overstates spend fivefold Files Needing Attention: model_prices_and_context_window.json, litellm/model_prices_and_context_window_backup.json
|
| Filename | Overview |
|---|---|
| model_prices_and_context_window.json | Adds Mercury 2.5 cache pricing and capability metadata, but currently overstates cached-token costs and lacks regression coverage |
| litellm/model_prices_and_context_window_backup.json | Mirrors the root cost-map changes, including the same premature standard cached-input rate |
Reviews (1): Last reviewed commit: "fix(cost map): add cached-input pricing ..." | Re-trigger Greptile
Codecov Report✅ All modified and coverable lines are covered by tests. 📢 Thoughts on this report? Let us know! |
Asserts the pricing/capability fields (including cache_read_input_token_cost and supports_prompt_caching) and that the root and backup cost maps stay in sync, following the existing per-model metadata test pattern.
|
Thanks for the review! Addressed the feedback: Test (P2): Added Cached rate (P1): The Note on CI: the failing |
|
The failing The only failure is: This is an async-mock issue in the auth suite, and it fails identically on the run before any test was added here. This PR only touches A re-run should be green once the flaky auth test is addressed on |
|
Superseded by #41112, which adds the same mercury-2.5 cache pricing verified against the Inception docs. |
#41016 Co-Authored-By: Nanduu24 <reachnanduu24@gmail.com> Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
Summary
inception/mercury-2.5is present in the cost map with input and output pricing, but is missing cached-input pricing and prompt-caching support:cache_read_input_token_cost— not setsupports_prompt_caching— not setAs a result, cached-token usage on Mercury 2.5 is not cost-tracked (silent inaccuracy), even though Inception offers cached input and the sibling model
inception/mercury-2already carries both fields.Changes
Adds to
inception/mercury-2.5in both the root cost map and thelitellm/backup copy (kept in sync):2e-08= $0.02 / 1M tokens, the standard cached-input rate (the current $0.004 is a temporary launch promo).Relevant issues
Addresses the cached-pricing portion of #40746 (the model entry itself was already added; these fields were still missing).
Type
🐛 Bug Fix / 💰 cost tracking
Testing