Skip to content

tests: use Qwen 3.5 9B for Together AI - #1385

Merged
HareeshBahuleyan merged 1 commit into
mainfrom
fix/together-agent-loop-model
Sep 9, 2026
Merged

HareeshBahuleyan merged 1 commit into
mainfrom
fix/together-agent-loop-model

Conversation

@HareeshBahuleyan

@HareeshBahuleyan HareeshBahuleyan commented Sep 9, 2026 •

Copy link
Copy Markdown
Contributor

Description

Why

Together's agent-loop integration test intermittently returns HTTP 400 input validation errors with openai/gpt-oss-20b. Together documents Qwen 3.5 9B for agentic multi-step function calling.

What changed

Use Qwen/Qwen3.5-9B for Together's general integration tests while keeping GPT-OSS for reasoning coverage.

Notes

  • uv run pytest tests/unit -q: 2,447 passed, 69 skipped.
  • Hooks for tests/conftest.py pass. Full pre-commit is blocked locally by an httpx versus httpx2 mypy type conflict in unchanged tests/unit/providers/test_openai_exceptions.py:188.
  • Targeted live Together CI run: test_agent_loop_sequential_tool_calls[together] passed on its first attempt in 15.93 seconds (run 34337514731).

PR Type

  • 🐛 Bug Fix

Relevant issues

CI run: https://github.com/mozilla-ai/any-llm/actions/runs/34332374589

Checklist

  • I understand the code I am submitting.
  • I have added unit tests that prove my fix/feature works
  • I have run this code locally and verified it fixes the issue.
  • New and existing tests pass locally
  • Documentation was updated where necessary
  • I have read and followed the contribution guidelines
  • AI Usage:
    • No AI was used.
    • AI was used for drafting/refactoring.
    • This is fully AI-generated.

AI Usage Information

  • AI Model used: OpenAI GPT-5.6 Sol
  • AI Developer Tool used: Pi
  • Any other info you'd like to share: The model selection was verified against Together's official model catalog and agentic function-calling documentation.

When answering questions by the reviewer, please respond yourself, do not copy/paste the reviewer comments into an AI system and paste back its answer. We want to discuss with you, not my AI :)

  • I am an AI Agent filling out this form (check box if true)

Summary by CodeRabbit

  • Tests
    • Updated the test configuration to use the Qwen 3.5 9B model for Together AI provider coverage.

@HareeshBahuleyan
HareeshBahuleyan deployed to integration-tests September 9, 2026 09:52 — with GitHub Actions Active
@coderabbitai

coderabbitai Bot commented Sep 9, 2026 •

Copy link
Copy Markdown

Review Change StackReview Change Stack

No actionable comments were generated in the recent review. 🎉

ℹ️ Recent review info
⚙️ Run configuration

Configuration used: Organization UI

Review profile: ASSERTIVE

Plan: Advanced

Run ID: 6944a38d-c4bb-43b3-9aa5-7ce7157238c5

📥 Commits

Reviewing files that changed from the base of the PR and between 7557ca5 and bce0087.

📒 Files selected for processing (1)
  • tests/conftest.py

Included review availability: Your plan provides up to 8 included reviews per hour; 6 remain after this review.


Walkthrough

Changes

Together test model

Layer / File(s) Summary
Update Together model mapping
tests/conftest.py
The Together provider test model changes from openai/gpt-oss-20b to Qwen/Qwen3.5-9B.

Suggested reviewers: njbrake

Priority: ⬇️ Low

Merge Risk: ⚪ Minimal · up to bce00

Together integration tests will use Qwen/Qwen3.5-9B for general coverage while GPT-OSS remains available for reasoning coverage. No actionable merge-blocking risk is identified.

🚥 Pre-merge checks | ✅ 4 | ❌ 1

❌ Failed checks (1 warning)

Check name Status Explanation Resolution
Docstring Coverage ⚠️ Warning Docstring coverage is 0.00% which is insufficient. The required threshold is 80.00%. Docstring coverage is scoped to functions touched by this diff. Analyzed 1 functions across 1 files. Write docstrings for the functions missing them to satisfy the coverage threshold.
✅ Passed checks (4 passed)
Check name Status Explanation
Linked Issues check ✅ Passed Check skipped because no linked issues were found for this pull request.
Out of Scope Changes check ✅ Passed Check skipped because no linked issues were found for this pull request.
Title check ✅ Passed The title is concise, specific, and accurately describes the model change for Together AI tests.
Description check ✅ Passed The description follows the required template, explains the reason and scope of the change, identifies the PR type, records relevant issues, documents validation results and limitations, and completes…
  • Fix all pre-merge checks with AI
✨ Finishing Touches 💡 1
📝 Generate docstrings 💡
  • Create stacked PR
  • Commit on current branch
🧪 Generate unit tests (beta)
  • Create PR with unit tests
  • Commit unit tests in branch fix/together-agent-loop-model

Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out.

❤️ Share

Comment @coderabbitai help to get the list of available commands.

@codecov

codecov Bot commented Sep 9, 2026

Copy link
Copy Markdown

Codecov Report

✅ All modified and coverable lines are covered by tests.
see 31 files with indirect coverage changes

🚀 New features to boost your workflow:
  • ❄️ Test Analytics: Detect flaky tests, report on failures, and find test suite problems.

@HareeshBahuleyan HareeshBahuleyan changed the title tests: use Qwen 3.5 9B for Together tests: use Qwen 3.5 9B for Together AI Sep 9, 2026
@HareeshBahuleyan HareeshBahuleyan self-assigned this Sep 9, 2026
@HareeshBahuleyan
HareeshBahuleyan merged commit d137811 into main Sep 9, 2026
27 checks passed
@HareeshBahuleyan
HareeshBahuleyan deleted the fix/together-agent-loop-model branch September 9, 2026 09:59
@github-actions github-actions Bot added the 1.27.2 Included in release 1.27.2 label Sep 10, 2026

This branch was successfully deployed

1 active deployment
integration-tests — bce0087c Deployed Sep 9, 2026 by HareeshBahuleyan via run-docs-tests #2896
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

1.27.2 Included in release 1.27.2

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant