llama.cpp
Ggml/cuda snake fusion hardening
#22912
Merged

Ggml/cuda snake fusion hardening #22912

ServeurpersoCom
ServeurpersoCom cuda: tighten snake fusion type checks for all operands (defensive, s…
e4f8ba35
ServeurpersoCom cuda: reject snake fusion when ne[2] or ne[3] > 1 (mirror vulkan PR r…
742dda0e
ServeurpersoCom ServeurpersoCom requested a review 148 days ago
am17an
am17an commented on 2026-05-10
ServeurpersoCom cuda: merge type_ok and types_ok into a single types_ok (address am17…
4cd7e1de
github-actions github-actions added Nvidia GPU
github-actions github-actions added ggml
am17an
am17an approved these changes on 2026-05-11
ServeurpersoCom
JohannesGaessler
JohannesGaessler approved these changes on 2026-05-11
am17an
ServeurpersoCom cuda: filter ADD/SUB/MUL/DIV in supports_op to F32/F16
5d325172
ORippler
ServeurpersoCom test-backend-ops: extend snake_fuse to rank-4 with ne[2]/ne[3] > 1 cases
c7d1a687
ServeurpersoCom ServeurpersoCom requested a review from ggerganov ggerganov 147 days ago
ServeurpersoCom
github-actions github-actions added testing
ServeurpersoCom
pwilkin
pwilkin approved these changes on 2026-05-11
ServeurpersoCom ServeurpersoCom merged e9366607 into master 147 days ago

Login to write a write a comment.

Login via GitHub

Assignees
No one assigned
Labels
Milestone