llama.cpp
0ac4b180
- qwen4exp: support a quantized KV cache in the QSA attention path
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
5 days ago
qwen4exp: support a quantized KV cache in the QSA attention path (cherry picked from commit 4c30574f81dc1115d08078c47b6cf8c789c0a842)
References
#27742 - model: add Qwen3.8-Flash-Next (qwen4exp)
Author
danielhanchen
Committer
danielhanchen
Parents
d4a943f9
Loading