Skip to content

opencl: q8_0 gemv precision improvement - #24923

Merged
lhez merged 1 commit into
ggml-org:masterfrom
qualcomm:sg/q8_0_precision_fix
Jun 23, 2026
Merged

lhez merged 1 commit into
ggml-org:masterfrom
qualcomm:sg/q8_0_precision_fix

Conversation

@shawngu-quic

Copy link
Copy Markdown
Contributor

Overview

Improve q8_0 gemv precision for some sensitive models.

Requirements

@shawngu-quic
shawngu-quic requested a review from a team as a code owner June 22, 2026 23:03
@github-actions github-actions Bot added ggml changes relating to the ggml tensor library for machine learning OpenCL Issues specific to the OpenCL backend labels Jun 22, 2026
@taronaeo taronaeo changed the title q8_0 gemv precision improvement opencl: q8_0 gemv precision improvement Jun 23, 2026
@lhez
lhez requested a review from max-krasnyansky June 23, 2026 04:02
@lhez
lhez merged commit 23ee879 into ggml-org:master Jun 23, 2026
4 checks passed
adrianhoehne pushed a commit to adrianhoehne/llama.cpp that referenced this pull request Jul 5, 2026
zommiommy pushed a commit to zommiommy/llama.cpp that referenced this pull request Aug 18, 2026
zbrad pushed a commit to zbrad/llama.cpp that referenced this pull request Sep 10, 2026
frostyautumnleaf pushed a commit to frostyautumnleaf/llama.cpp that referenced this pull request Oct 5, 2026
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

ggml changes relating to the ggml tensor library for machine learning OpenCL Issues specific to the OpenCL backend

Projects

None yet

Development

Successfully merging this pull request may close these issues.

3 participants