llama.cpp
CUDA: faster non k-quant mul_mat_q kernels
#2483
Merged
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Overview
Commits
1
Changes
View On
GitHub
CUDA: faster non k-quant mul_mat_q kernels
#2483
JohannesGaessler
merged 1 commit into
ggml-org:master
from
JohannesGaessler:cuda-faster-mmq-3
slaren
approved these changes on 2023-08-02
CUDA: faster non k-quant mul_mat_q kernels
d6154f5b
JohannesGaessler
force pushed
to
d6154f5b
2 years ago
JohannesGaessler
merged
468ea24f
into master
2 years ago
Login to write a write a comment.
Login via GitHub
Reviewers
slaren
Assignees
No one assigned
Labels
None yet
Milestone
No milestone
Login to write a write a comment.
Login via GitHub