llama.cpp
311d4211
- memory : avoid allocating V cache for indexer (it's not used) (#28330)
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
13 days ago
memory : avoid allocating V cache for indexer (it's not used) (#28330) Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com>
References
#28330 - memory : avoid allocating V cache for indexer (it's not used) in Qwen3.8-Flash-Next (qwen4exp)
Author
fairydreaming
Parents
72797e89
Loading