Skip to content
Navigation Menu
Sign in
Appearance settings
Platform
AI CODE CREATION
GitHub Copilot
Write better code with AI
GitHub Copilot app
Direct agents from issue to merge
MCP Registry
Integrate external tools
DEVELOPER WORKFLOWS
Actions
Automate any workflow
Codespaces
Instant dev environments
Issues
Plan and track work
Code Review
Manage code changes
Code Quality
Enforce quality at merge
APPLICATION SECURITY
GitHub Advanced Security
Find and fix vulnerabilities
Code security
Secure your code as you build
Secret protection
Stop leaks before they start
EXPLORE
Why GitHub
Documentation
Blog
Changelog
Marketplace
View all features
Solutions
BY COMPANY SIZE
Enterprises
Small and medium teams
Startups
Nonprofits
BY USE CASE
App Modernization
DevSecOps
DevOps
CI/CD
View all use cases
BY INDUSTRY
Healthcare
Financial services
Manufacturing
Government
View all industries
View all solutions
Resources
EXPLORE BY TOPIC
AI
Software Development
DevOps
Security
View all topics
EXPLORE BY TYPE
Customer stories
Events & webinars
Ebooks & reports
Business insights
GitHub Skills
SUPPORT & SERVICES
Documentation
Customer support
Community forum
Trust center
Partners
View all resources
Open Source
COMMUNITY
GitHub Sponsors
Fund open source developers
PROGRAMS
Security Lab
Maintainer Community
GitHub Stars
Archive Program
REPOSITORIES
Topics
Trending
Collections
Enterprise
ENTERPRISE SOLUTIONS
Enterprise platform
AI-powered developer platform
AVAILABLE ADD-ONS
GitHub Advanced Security
Enterprise-grade security features
Copilot for Business
Enterprise-grade AI features
Premium Support
Enterprise-grade 24/7 support
Pricing
Search
/
Sign in
Sign up
Appearance settings
You signed in with another tab or window.
Reload
to refresh your session.
You signed out in another tab or window.
Reload
to refresh your session.
You switched accounts on another tab or window.
Reload
to refresh your session.
Dismiss alert
{{ message }}
Uh oh!
There was an error while loading.
Please reload this page
.
NVIDIA
/
TensorRT-LLM
Public
Notifications
You must be signed in to change notification settings
Fork
2.8k
Star
14.6k
Code
Issues
597
Pull requests
896
Discussions
Actions
Projects
Security and quality
0
Insights
Additional navigation options
Code
Issues
Pull requests
Discussions
Actions
Projects
Security and quality
Insights
[TRTLLM-15316][feat] sm107 gemm + quant
- #17485
#17485
Merged
BowenFu
merged 13 commits into
NVIDIA:main
NVIDIA/TensorRT-LLM:main
from
farazkh80:rubin_feat/gemm_quant
farazkh80/TensorRT-LLM:rubin_feat/gemm_quant
Copy head branch name to clipboard
Aug 29, 2026
Conversation
Commits
13
(13)
Checks
Files changed
Merged
[TRTLLM-15316][feat] sm107 gemm + quant
#17485
BowenFu
merged 13 commits into
NVIDIA:main
NVIDIA/TensorRT-LLM:main
from
farazkh80:rubin_feat/gemm_quant
farazkh80/TensorRT-LLM:rubin_feat/gemm_quant
Copy head branch name to clipboard
Commits
Commits on Aug 25, 2026
[None][feat] SM107: build and dispatch CUTLASS GEMM kernels for the SM100 family
Show description for d097df3
4 people
committed
d097df3
View commit details
Copy full SHA for d097df3
Browse repository at this point
[None][refactor] parameterize block-scale quantization on SF vector size
Show description for 1ffbedd
mingyangHao
and
farazkh80
committed
1ffbedd
View commit details
Copy full SHA for 1ffbedd
Browse repository at this point
[None][feat] SM107 support in the FP8 quantize and scaled-MM torch ops
Show description for c26802f
3 people
committed
c26802f
View commit details
Copy full SHA for c26802f
Browse repository at this point
[None][fix] address review: R128c4 quantize tests + MxFP8 template default
Show description for 9f29f40
mingyangHao
authored and
farazkh80
committed
9f29f40
View commit details
Copy full SHA for 9f29f40
Browse repository at this point
[None][fix] address review: keep legacy SF layout default, restrict SM107 to 100f kernels
Show description for 73a3615
farazkh80
committed
73a3615
View commit details
Copy full SHA for 73a3615
Browse repository at this point
[None][test] run the R128c4 quantize tests on the full SM100 family
Show description for 53a5357
farazkh80
committed
53a5357
View commit details
Copy full SHA for 53a5357
Browse repository at this point
[None][fix] address review: make gemm no-kernel fail-fast self-describing
Show description for 8290c61
farazkh80
committed
8290c61
View commit details
Copy full SHA for 8290c61
Browse repository at this point
Commits on Aug 26, 2026
Merge branch 'main' into rubin_feat/gemm_quant
BowenFu
authored
640f481
View commit details
Copy full SHA for 640f481
Browse repository at this point
Merge branch 'main' into rubin_feat/gemm_quant
BowenFu
authored
d83c3bb
View commit details
Copy full SHA for d83c3bb
Browse repository at this point
Merge remote-tracking branch 'github/main' into rubin_feat/gemm_quant
Show description for d8dd182
farazkh80
committed
d8dd182
View commit details
Copy full SHA for d8dd182
Browse repository at this point
Commits on Aug 27, 2026
Merge remote-tracking branch 'github/main' into rubin_feat/gemm_quant
Show description for c8ff146
farazkh80
committed
c8ff146
View commit details
Copy full SHA for c8ff146
Browse repository at this point
Merge remote-tracking branch 'github/main' into rubin_feat/gemm_quant
Show description for 27c8293
farazkh80
committed
27c8293
View commit details
Copy full SHA for 27c8293
Browse repository at this point
Commits on Aug 28, 2026
Merge branch 'main' into rubin_feat/gemm_quant
BowenFu
authored
be67661
View commit details
Copy full SHA for be67661
Browse repository at this point
You can’t perform that action at this time.