llama.cpp
llama : allocate GLM_DSA indexer cache only in "full" indexer layers
#26474
Merged

llama : allocate GLM_DSA indexer cache only in "full" indexer layers #26474

fairydreaming
sszymczy llama : allocate indexer cache only in "full" indexer layers
bda1f551
sszymczy Merge remote-tracking branch 'upstream/master' into glm-lid-cache-filter
32326ca9
fairydreaming fairydreaming marked this pull request as ready for review 16 days ago
fairydreaming fairydreaming requested a review from CISC CISC 16 days ago
fairydreaming fairydreaming requested a review from ggerganov ggerganov 16 days ago
CISC
CISC approved these changes on 2026-08-03
ggerganov
ggerganov approved these changes on 2026-08-03
fairydreaming fairydreaming merged 563dec81 into master 15 days ago

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone