llama.cpp
CUDA: fix thread/block count in quantized cpy kernel launches
#26731
Merged
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Overview
Commits
2
Changes
View On
GitHub
Commits
CUDA: fix thread/block count in quantized cpy kernel launches
grafail
committed
10 days ago
tests: add uneven block count cpy case
grafail
committed
10 days ago
Loading