llama.cpp
vulkan: fix stale prealloc_y reuse across flash attention and soft_max
#29591
Merged

vulkan: fix stale prealloc_y reuse across flash attention and soft_max #29591

fxgsell
fxgsell fxgsell requested a review from ggerganov ggerganov 10 days ago
fxgsell fxgsell requested a review 10 days ago
github-actions github-actions added testing
github-actions github-actions added Vulkan
github-actions github-actions added ggml
jeffbolznv
fxgsell vulkan: fix stale prealloc_y reuse across flash attention and soft_max
5a20ad1a
fxgsell fxgsell force pushed from 3ddd26a1 to 5a20ad1a 10 days ago
fxgsell
jeffbolznv
jeffbolznv approved these changes on 2026-09-28
jeffbolznv
jeffbolznv
jeffbolznv approved these changes on 2026-10-01
0cc4m
0cc4m approved these changes on 2026-10-05
0cc4m 0cc4m merged 806eee98 into master 4 days ago

Login to write a write a comment.

Login via GitHub

Assignees
No one assigned
Labels
Milestone