llama.cpp
69bf6437 - CUDA: fix thread/block count in quantized cpy kernel launches (#26731)

Commit
7 days ago
CUDA: fix thread/block count in quantized cpy kernel launches (#26731) * CUDA: fix thread/block count in quantized cpy kernel launches * tests: add uneven block count cpy case
Author
Parents
Loading