llama.cpp
c350a40b
- Performance tune for gemma4-26b-a4b flash attention shape. (#28450)
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
5 days ago
Performance tune for gemma4-26b-a4b flash attention shape. (#28450)
References
#28450 - Performance tune for gemma4-26b-a4b flash attention shape.
Author
frobnitzem
Parents
9b421fa9
Loading