Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
NVIDIA
/
TensorRT-LLM
Public
Notifications
You must be signed in to change notification settings
Fork
2.7k
Star
14.6k
Code
Issues
577
Pull requests
870
Discussions
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Security and quality
Insights
[None][fix] Unwaive supported attention backend cases
- #15838
#15838
Merged
yuxianq
merged 13 commits into
NVIDIA:main
NVIDIA/TensorRT-LLM:main
from
yuxianq:unwaive-attn-backend-test-cases
yuxianq/TensorRT-LLM:unwaive-attn-backend-test-cases
Copy head branch name to clipboard
Jul 8, 2026
Conversation
Commits
13
(13)
Checks
Files changed
Merged
[None][fix] Unwaive supported attention backend cases
#15838
yuxianq
merged 13 commits into
NVIDIA:main
NVIDIA/TensorRT-LLM:main
from
yuxianq:unwaive-attn-backend-test-cases
yuxianq/TensorRT-LLM:unwaive-attn-backend-test-cases
Copy head branch name to clipboard
Commits
Commits on Jul 7, 2026
[None][fix] enable TRTLLM MLA context backend tests
Show description for bb93032
yuxianq
committed
bb93032
View commit details
Copy full SHA for bb93032
Browse repository at this point
[None][fix] enable TRTLLM Blackwell MLA generation
Show description for ccf8516
yuxianq
committed
ccf8516
View commit details
Copy full SHA for ccf8516
Browse repository at this point
[None][fix] reduce TRTLLM attention backend exclusions
Show description for 3e2a315
yuxianq
committed
3e2a315
View commit details
Copy full SHA for 3e2a315
Browse repository at this point
[None][fix] unwaive stable attention backend cases
Show description for ca66d0c
yuxianq
committed
ca66d0c
View commit details
Copy full SHA for ca66d0c
Browse repository at this point
[None][fix] remove obsolete Hopper FMHA workaround
Show description for 4f259ca
yuxianq
committed
4f259ca
View commit details
Copy full SHA for 4f259ca
Browse repository at this point
[None][fix] support large FlashInfer KV appends
Show description for 52d1b18
yuxianq
committed
52d1b18
View commit details
Copy full SHA for 52d1b18
Browse repository at this point
[None][fix] restore missing MLA generation kernel guard
Show description for 96f9813
yuxianq
committed
96f9813
View commit details
Copy full SHA for 96f9813
Browse repository at this point
[None][fix] support FP8 KV MLA in TRTLLM-gen
Show description for aadd5a4
yuxianq
committed
aadd5a4
View commit details
Copy full SHA for aadd5a4
Browse repository at this point
[None][fix] fall back for FP8 multi-token MLA
Show description for ec49210
yuxianq
committed
ec49210
View commit details
Copy full SHA for ec49210
Browse repository at this point
[None][fix] gate FP8 multi-token MLA fallback by FlashInfer version
Show description for f497c9c
yuxianq
committed
f497c9c
View commit details
Copy full SHA for f497c9c
Browse repository at this point
[None][fix] remove obsolete FlashInfer MLA fallback
Show description for a385754
yuxianq
committed
a385754
View commit details
Copy full SHA for a385754
Browse repository at this point
[None][fix] fix FP8 KV-only attention fallback
Show description for 3a40842
yuxianq
committed
3a40842
View commit details
Copy full SHA for 3a40842
Browse repository at this point
[None][fix] avoid forcing paged context FMHA for FP8 KV
Show description for 7483843
yuxianq
committed
7483843
View commit details
Copy full SHA for 7483843
Browse repository at this point
You can’t perform that action at this time.