llama : allocate GLM_DSA indexer cache only in "full" indexer layers #26474
llama : allocate indexer cache only in "full" indexer layers
bda1f551
Merge remote-tracking branch 'upstream/master' into glm-lid-cache-filter
32326ca9
fairydreaming
marked this pull request as ready for review 16 days ago
CISC
approved these changes
on 2026-08-03
ggerganov
approved these changes
on 2026-08-03
Assignees
No one assigned
Login to write a write a comment.
Login via GitHub