llama.cpp
ggml-webgpu: improve i-quants mul_mat performance and speed up prefill
#24530
Merged

ggml-webgpu: improve i-quants mul_mat performance and speed up prefill #24530

yomaytk
yomaytk Improve prefill speeds for i-quants
5fb0c144
yomaytk yomaytk requested a review 42 days ago
github-actions github-actions added ggml
github-actions github-actions added WebGPU
CISC
CISC commented on 2026-06-12
yomaytk Fix #if defined() usage in preprocessor guards.
1060fdc3
CISC
CISC approved these changes on 2026-06-13
reeselevine
reeselevine approved these changes on 2026-06-15
reeselevine reeselevine merged 6e9007ae into master 39 days ago
yomaytk yomaytk deleted the improve-mul_mat-iquants branch 39 days ago

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone