llama.cpp
opencl: Q6_K GEMM/GEMV fix for ne01 of weights that are not multiples of 128.
#25464
Merged

opencl: Q6_K GEMM/GEMV fix for ne01 of weights that are not multiples of 128. #25464

wanghqc
wanghqc opencl: fix garbled output for Q6_K weights with ne01 % 128 != 0 on A…
3ff664d2
wanghqc opencl: reserve alignment slack for the SOA subbuffer carve in alloc …
f7eda48a
lhez opencl: use lm based q6_k mm when ne1 is not multiple of 128
0f2bce23
wanghqc wanghqc requested a review 55 days ago
github-actions github-actions added ggml
github-actions github-actions added OpenCL
lhez
lhez approved these changes on 2026-07-08
lhez lhez requested a review from max-krasnyansky max-krasnyansky 55 days ago
max-krasnyansky
max-krasnyansky approved these changes on 2026-07-08
max-krasnyansky max-krasnyansky merged 92366df3 into master 54 days ago

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone