llama.cpp
opencl: use a better matmul path on two Adreno GPU generations
#27640
Merged

opencl: use a better matmul path on two Adreno GPU generations #27640

wanghqc
wanghqc opencl: default the Adreno xmem F16xF32 GEMM on for X2E
a26f7261
wanghqc opencl: bypass the tiled f32 GEMM on the Adreno A7X
7b68a59c
github-actions github-actions added ggml
github-actions github-actions added OpenCL
lhez opencl: enable xmem GEMM for adreno by default
84b78229
lhez lhez marked this pull request as ready for review 8 days ago
lhez lhez requested a review 8 days ago
lhez
lhez approved these changes on 2026-08-29
lhez lhez requested a review from max-krasnyansky max-krasnyansky 8 days ago
max-krasnyansky
max-krasnyansky approved these changes on 2026-08-29
max-krasnyansky max-krasnyansky merged c841aeeb into master 7 days ago

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone