Skip to content

[Bugfix] Fix KV cache sizing and allocation for hybrid Mamba/attention models - #37429

Open
swtb3 wants to merge 4 commits into
vllm-project:mainfrom
swtb3:fix/hybrid-mamba-compact-allocation
Open

[Bugfix] Fix KV cache sizing and allocation for hybrid Mamba/attention models#37429
swtb3 wants to merge 4 commits into
vllm-project:mainfrom
swtb3:fix/hybrid-mamba-compact-allocation

Merge branch 'main' into fix/hybrid-mamba-compact-allocation

f74c1cb
Select commit
Loading
Failed to load commit list.
DCO / DCO succeeded Mar 19, 2026 in 1s

DCO

All commits are signed off!