llama.cpp
memory : avoid allocating V cache for indexer (it's not used) in Qwen3.8-Flash-Next (qwen4exp)
#28330
Open
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Overview
Commits
1
Changes
View On
GitHub
memory : avoid allocating V cache for indexer (it's not used) in Qwen3.8-Flash-Next (qwen4exp)
#28330
fairydreaming
wants to merge 1 commit into
ggml-org:master
from
fairydreaming:qwen4exp-no-indexer-v-cache
memory : avoid allocating V cache for indexer (it's not used)
b12a411b
fairydreaming
requested a review
from
ggerganov
7 hours ago
pwilkin
approved these changes on 2026-09-03
Login to write a write a comment.
Login via GitHub
Reviewers
pwilkin
ggerganov
Assignees
No one assigned
Labels
None yet
Milestone
No milestone
Login to write a write a comment.
Login via GitHub