llama.cpp
cuda: guard the iq4_nl dequantize row kernel against short rows
#29683
Merged

cuda: guard the iq4_nl dequantize row kernel against short rows #29683

yeahdongcn
yeahdongcn cuda: guard the iq4_nl dequantize row kernel against short rows
908d76a7
github-actions github-actions added ggml
github-actions github-actions added CUDA
yeahdongcn yeahdongcn marked this pull request as ready for review 10 days ago
yeahdongcn yeahdongcn requested a review 10 days ago
yeahdongcn yeahdongcn requested a review from CISC CISC 10 days ago
yeahdongcn yeahdongcn requested a review from slaren slaren 10 days ago
yeahdongcn yeahdongcn requested a review from JohannesGaessler JohannesGaessler 10 days ago
CISC CISC removed review request from slaren slaren 10 days ago
am17an
am17an approved these changes on 2026-09-30
JohannesGaessler
JohannesGaessler approved these changes on 2026-09-30
JohannesGaessler JohannesGaessler merged f872b591 into master 9 days ago

Login to write a write a comment.

Login via GitHub

Assignees
No one assigned
Labels
Milestone