llama.cpp
Fix data race in CUDA's "cpy" kernel (influences GGML's DUP, CONT operations).
#20507
Merged

Fix data race in CUDA's "cpy" kernel (influences GGML's DUP, CONT operations). #20507

Exile333
Exile333 Fix datarace in CUDA's "cpy" kernel.
47c1c71e
ggerganov ggerganov requested a review 136 days ago
ggerganov ggerganov removed review request 136 days ago
github-actions github-actions added Nvidia GPU
github-actions github-actions added ggml
JohannesGaessler
JohannesGaessler approved these changes on 2026-03-13
ORippler
ORippler commented on 2026-03-13
Exile333
Exile333 Remove extra barrier by using more of shared memory.
79c9996f
Exile333
Exile333
JohannesGaessler
JohannesGaessler approved these changes on 2026-03-13
am17an am17an merged 5a32a9b8 into master 135 days ago
Exile333 Exile333 deleted the cuda_cpy_datarace_fix branch 133 days ago

Login to write a write a comment.

Login via GitHub

Assignees
No one assigned
Labels
Milestone