Skip to content

feat(models): sync together_ai model registry - #39079

Closed
mateo-berri wants to merge 1 commit into
litellm_internal_stagingfrom
litellm_together_registry_sync_2026-09-01
Closed

feat(models): sync together_ai model registry#39079
mateo-berri wants to merge 1 commit into
litellm_internal_stagingfrom
litellm_together_registry_sync_2026-09-01

Conversation

@mateo-berri

Copy link
Copy Markdown
Contributor

Automated daily sync of the together_ai entries in model_prices_and_context_window.json against GET https://api.together.ai/v1/models?serverless and https://docs.together.ai/docs/deprecations.md by scripts/sync_together_ai_models.py.

Added (1)

  • together_ai/Qwen/Qwen3.8-Flash

Updated (0)

  • none

Marked deprecated (0)

  • none

Returned to the catalog (0)

  • none

Warnings needing a human call (8)

  • together_ai/Qwen/Qwen3.8-Flash added without a capability rule; review its tools/vision/reasoning support and add one
  • together_ai/BAAI/bge-base-en-v1.5 is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call
  • together_ai/Qwen/Qwen2.5-7B-Instruct-Turbo is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call
  • together_ai/baai/bge-base-en-v1.5 is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call
  • together_ai/deepseek-ai/DeepSeek-V3 is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call
  • together_ai/moonshotai/Kimi-K2-Instruct is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call
  • together_ai/togethercomputer/CodeLlama-34b-Instruct is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call
  • together_ai/zai-org/GLM-4.6 is absent from the serverless catalog with no removal date in the docs; needs a human deprecation call

Catalog model types outside the sync's token-pricing scope, skipped: audio (5), image (29), transcribe (4), video (38)

@CLAassistant

Copy link
Copy Markdown

CLA assistant check
Thank you for your submission! We really appreciate it. Like many open source projects, we ask that you sign our Contributor License Agreement before we can accept your contribution.
You have signed the CLA already but the status is still pending? Let us recheck it.

@greptile-apps

greptile-apps Bot commented Sep 1, 2026

Copy link
Copy Markdown
Contributor

Greptile Summary

This PR synchronizes the Together AI model registry by adding the Qwen3.8-Flash chat model to both canonical pricing files.

  • Adds input and output token pricing for together_ai/Qwen/Qwen3.8-Flash.
  • Records a one-million-token input/context limit.
  • Keeps the primary and backup model registries identical.

Confidence Score: 5/5

The PR appears safe to merge based on the reviewed registry changes.

The new entry is schema-conformant and synchronized across both registry files, and no concrete reachable failure was established from the changed metadata.

Important Files Changed

Filename Overview
model_prices_and_context_window.json Adds a schema-conformant Together AI model entry; no concrete changed-code defect was established.
litellm/model_prices_and_context_window_backup.json Mirrors the new Together AI entry from the primary registry and preserves backup synchronization.

Reviews (1): Last reviewed commit: "feat(models): sync together_ai model reg..." | Re-trigger Greptile

@codspeed-hq

codspeed-hq Bot commented Sep 1, 2026

Copy link
Copy Markdown
Contributor

Merging this PR will not alter performance

✅ 31 untouched benchmarks


Comparing litellm_together_registry_sync_2026-09-01 (00807d9) with litellm_internal_staging (ec3f818)

Open in CodSpeed

@codecov

codecov Bot commented Sep 1, 2026

Copy link
Copy Markdown

Codecov Report

✅ All modified and coverable lines are covered by tests.

📢 Thoughts on this report? Let us know!

@devin-ai-integration

Copy link
Copy Markdown
Contributor

Closing: together_ai/Qwen/Qwen3.8-Flash already landed on litellm_internal_staging via #38990. Today's rolling registry PR is #39170.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants