Skip to content

[bugfix] skip conch kernel for g_idx reordering - #45072

Merged
tjtanaa merged 1 commit into
vllm-project:mainfrom
ROCm:fix_conchkernel_selection
Jun 10, 2026
Merged

[bugfix] skip conch kernel for g_idx reordering#45072
tjtanaa merged 1 commit into
vllm-project:mainfrom
ROCm:fix_conchkernel_selection

Conversation

@divakar-amd

@divakar-amd divakar-amd commented Jun 9, 2026

Copy link
Copy Markdown
Contributor

Conch kernel doesn't support g_idx reordering.

The issue is not ROCm specific. CUDA fails to capture this because the kernel ordering on cuda picks Marlin kernel before conch which handles g_idx reordering.
Kernel ordering preference for ROCm & CUDA can be referred here - LINK

Resolves the following test: lora/test_quant_model.py::test_quant_model_lora[model0]. This PR will safely skip selecting the conchkernel and on rocm the next kernel (ExllamaLinearKernel) in the preference order will be chosen.

Signed-off-by: Divakar Verma <divakar.verma@amd.com>
@mergify mergify Bot added the bug Something isn't working label Jun 9, 2026
@AndreasKaratzas AndreasKaratzas added the ready ONLY add when PR is ready to merge/full CI is needed label Jun 9, 2026
@tjtanaa
tjtanaa merged commit 166d14e into vllm-project:main Jun 10, 2026
66 checks passed
@AndreasKaratzas
AndreasKaratzas deleted the fix_conchkernel_selection branch June 10, 2026 16:02
wcynb1023 pushed a commit to wcynb1023/vllm that referenced this pull request Jun 11, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
Saddss pushed a commit to Saddss/vllm that referenced this pull request Jun 14, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
vivek8123 pushed a commit to odh-on-pz/vllm-upstream that referenced this pull request Jun 18, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
divineearthly pushed a commit to divineearthly/vllm that referenced this pull request Jun 19, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
Signed-off-by: divineearthly <divineearthly@gmail.com>
nkzhenhua pushed a commit to nkzhenhua/vllm that referenced this pull request Jun 24, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
Dao007forever pushed a commit to Dao007forever/vllm that referenced this pull request Jul 18, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
philippesic pushed a commit to philippesic/vllm-semantic-cache that referenced this pull request Jul 19, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
plasticchris pushed a commit to plasticchris/vllm that referenced this pull request Jul 20, 2026
Signed-off-by: Divakar Verma <divakar.verma@amd.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

bug Something isn't working ready ONLY add when PR is ready to merge/full CI is needed

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants