llama.cpp
ggml-webgpu: Fix bug in FlashAttention support check
#22492
Merged
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Overview
Commits
2
Changes
View On
GitHub
Commits
Fix flashattention support check for devices that don't support subgroups
reeselevine
committed
98 days ago
set path to none if kv_tile doesn't fit
reeselevine
committed
98 days ago
Loading