llama.cpp
ggml-webgpu: Fix bug in FlashAttention support check
#22492
Merged

Commits
  • Fix flashattention support check for devices that don't support subgroups
    reeselevine committed 98 days ago
  • set path to none if kv_tile doesn't fit
    reeselevine committed 98 days ago
Loading