llama.cpp
563dec81
- llama : allocate indexer cache only in "full" indexer layers (#26474)
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
13 days ago
llama : allocate indexer cache only in "full" indexer layers (#26474) Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com>
References
#26474 - llama : allocate GLM_DSA indexer cache only in "full" indexer layers
Author
fairydreaming
Parents
96278e39
Loading