llama.cpp
78d2f524
- cuda : concat implementation for quantized types (#25303)
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
48 days ago
cuda : concat implementation for quantized types (#25303) * cuda : concat implementation for quantized types * chore : apply am17an clever suggestion to shorten the code --------- Co-authored-by: Stanisław Szymczyk <sszymczy@gmail.com>
References
#25303 - cuda : concat implementation for quantized types
Author
fairydreaming
Parents
a4107133
Loading