Skip to content

opencl: add bin kernels kernel_gemm_moe_q4_0_q8_1_dp4a_bin, kernel_gemm_moe_mxfp4_q8_1_dp4a_bin - #27768

Merged
lhez merged 1 commit into
ggml-org:masterfrom
qualcomm:sg/moe-dp4a-bin-kernel-upstream
Aug 27, 2026
Merged

opencl: add bin kernels kernel_gemm_moe_q4_0_q8_1_dp4a_bin, kernel_gemm_moe_mxfp4_q8_1_dp4a_bin#27768
lhez merged 1 commit into
ggml-org:masterfrom
qualcomm:sg/moe-dp4a-bin-kernel-upstream

Conversation

@shawngu-quic

@shawngu-quic shawngu-quic commented Aug 26, 2026

Copy link
Copy Markdown
Contributor

Overview

Adreno optimized bin kernel support for MoE OpenCL kernels using dp4a. Including mxfp4 and Q4_0 in this commit.

Requirements

@shawngu-quic
shawngu-quic requested a review from a team as a code owner August 26, 2026 22:04
@github-actions github-actions Bot added ggml changes relating to the ggml tensor library for machine learning OpenCL Issues specific to the OpenCL backend labels Aug 26, 2026
@lhez
lhez requested a review from max-krasnyansky August 26, 2026 22:19
@lhez
lhez merged commit 5854625 into ggml-org:master Aug 27, 2026
24 of 27 checks passed
thecodacus pushed a commit to thecodacus/llama.cpp that referenced this pull request Sep 7, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

ggml changes relating to the ggml tensor library for machine learning OpenCL Issues specific to the OpenCL backend

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants