Repository navigation
feat(models): add Tencent Hy3 model - #3325
Conversation
|
No actionable comments were generated in the recent review. 🎉 ℹ️ Recent review info⚙️ Run configurationConfiguration used: Repository UI Review profile: CHILL Plan: Pro Plus Run ID: 📒 Files selected for processing (1)
WalkthroughAdds Tencent’s HY3 model with DeepInfra and Novita provider configurations, then registers the definitions in the shared models array. ChangesTencent model registry
Estimated code review effort: 2 (Simple) | ~10 minutes Possibly related PRs
Suggested reviewers: 🚥 Pre-merge checks | ✅ 5✅ Passed checks (5 passed)
✨ Finishing Touches🧪 Generate unit tests (beta)
Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out. Comment |
There was a problem hiding this comment.
Actionable comments posted: 1
🤖 Prompt for all review comments with AI agents
Verify each finding against current code. Fix only still-valid issues, skip the
rest with a brief reason, keep changes minimal, and validate.
Inline comments:
In `@packages/models/src/models/tencent.ts`:
- Around line 21-25: Add the shared reasoningEfforts values ["no_think", "low",
"high"] to both Hy3 model mappings alongside reasoning: true, and update the Hy3
request-building path to forward the selected reasoning effort as the
OpenAI-compatible reasoning_effort body field.
🪄 Autofix (Beta)
Fix all unresolved CodeRabbit comments on this PR:
- Push a commit to this branch (recommended)
- Create a new PR with the fixes
ℹ️ Review info
⚙️ Run configuration
Configuration used: Repository UI
Review profile: CHILL
Plan: Pro Plus
Run ID: a5ed714d-2cba-4e70-b992-2415bbc29939
⛔ Files ignored due to path filters (1)
pnpm-lock.yamlis excluded by!**/pnpm-lock.yaml
📒 Files selected for processing (2)
packages/models/src/models.tspackages/models/src/models/tencent.ts
301b369 to
7d6f859
Compare
20062a9 to
ac23f48
Compare
Add Tencent's Hy3 (295B Mixture-of-Experts, 21B active) with provider
mappings for DeepInfra and Novita. Native 256K context and three
reasoning modes for complex reasoning and agentic tasks.
Provider mappings (per-token prices in e-6 notation):
DeepInfra (tencent/Hy3): input 0.14e-6, cached 0.035e-6, output 0.58e-6
context 262144, maxOutput 131072, fp8, reasoning/tools/jsonOutput
Novita (tencent/hy3): input 0.14e-6, cached 0.035e-6, output 0.58e-6
context 262144, maxOutput 262144, reasoning/tools/jsonOutput
reasoningEfforts: ["none", "low", "high"] on both. Verified live against
DeepInfra's API that "none" disables reasoning and "low"/"high" enable it.
DeepInfra's maxOutput is 131072 (max_output_tokens from its model metadata);
the page's "max_tokens 262144" is the full context window, not the output cap.
Novita's max_output_tokens is 262144.
Refs:
https://deepinfra.com/tencent/Hy3
https://novita.ai/models/model-detail/tencent-hy3
https://openrouter.ai/tencent/hy3
https://github.com/Tencent-Hunyuan/Hy3
Co-Authored-By: Claude <noreply@anthropic.com>
ac23f48 to
0de5b76
Compare
Novita's tencent/hy3 rejects the json_object response format (400, "Supported formats: json_schema"); it only supports json_schema. DeepInfra's tencent/Hy3 supports both. Verified against both provider APIs directly and via the scoped e2e run. Co-Authored-By: Claude <noreply@anthropic.com>
|
Ran the scoped e2e suite for both mappings ( One real catalogue bug found and fixed (pushed as 2f0de7e): Novita's It does support
Both verified with direct provider calls, not just through the gateway. Results after the fix
The only remaining failure is Other metadata verified against the provider APIsPrices ($0.14 / $0.035 cached / $0.58 per M), context 262144, max output (131072 DeepInfra / 262144 Novita), |
Summary
Adds Tencent's Hy3 model to the catalogue with provider mappings for DeepInfra and NovitaAI.
Hy3 is Tencent's 295B Mixture-of-Experts model (21B active params) with native 256K context and three reasoning modes, targeting complex reasoning and agentic tasks.
Provider mappings
tencent/Hy3tencent/hy3none,low,high)none,low,high)reasoningEfforts: ["none", "low", "high"]on both. DeepInfra's max output is 131072 (max_output_tokensfrom its model metadata; the page'smax_tokens: 262144is the full context window, not the output cap). Novita'smax_output_tokensis 262144.References
🤖 Generated with Claude Code
Summary by CodeRabbit