llama.cpp
0ac4b180 - qwen4exp: support a quantized KV cache in the QSA attention path

Commit
5 days ago
qwen4exp: support a quantized KV cache in the QSA attention path (cherry picked from commit 4c30574f81dc1115d08078c47b6cf8c789c0a842)
Author
danielhanchen
Committer
Parents
Loading