Skip to content

opencl: add bin kernel kernel_gemm_noshuffle_q4_0_f32_32b_trans_ila_a8_bin - #28268

Merged
lhez merged 1 commit into
ggml-org:masterfrom
qualcomm:q4_0-a8-bin-kernel
Sep 10, 2026
Merged

lhez merged 1 commit into
ggml-org:masterfrom
qualcomm:q4_0-a8-bin-kernel

Conversation

@shaofeiqi

Copy link
Copy Markdown
Contributor

Overview

Add optimized Q4_0 non-MoE GEMM as bin kernel, plus the matching GEMV kernel for Adreno.

Requirements

@github-actions github-actions Bot added ggml changes relating to the ggml tensor library for machine learning OpenCL Issues specific to the OpenCL backend labels Sep 2, 2026
@shaofeiqi
shaofeiqi marked this pull request as ready for review September 10, 2026 05:54
@shaofeiqi
shaofeiqi requested a review from a team as a code owner September 10, 2026 05:54
@lhez
lhez merged commit df03399 into ggml-org:master Sep 10, 2026
26 of 27 checks passed
@BrewTestBot BrewTestBot mentioned this pull request Sep 14, 2026
1 task done
pl752 pushed a commit to pl752/llama.cpp that referenced this pull request Sep 15, 2026
zsogitbe pushed a commit to zsogitbe/llama.cpp that referenced this pull request Sep 17, 2026
frostyautumnleaf pushed a commit to frostyautumnleaf/llama.cpp that referenced this pull request Oct 5, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

ggml changes relating to the ggml tensor library for machine learning OpenCL Issues specific to the OpenCL backend

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants