opencl: Q6_K GEMM/GEMV fix for ne01 of weights that are not multiples of 128. #25464
opencl: fix garbled output for Q6_K weights with ne01 % 128 != 0 on A…
3ff664d2
opencl: reserve alignment slack for the SOA subbuffer carve in alloc …
f7eda48a
opencl: use lm based q6_k mm when ne1 is not multiple of 128
0f2bce23
wanghqc
requested a review
55 days ago
lhez
approved these changes
on 2026-07-08
Assignees
No one assigned
Login to write a write a comment.
Login via GitHub