llama.cpp
opencl: fold the gpt-oss MoE per-expert bias adds into the epilogue (op/kernel fusion)
#26431
Merged

opencl: fold the gpt-oss MoE per-expert bias adds into the epilogue (op/kernel fusion) #26431

wanghqc
github-actions github-actions added ggml
github-actions github-actions added OpenCL
wanghqc opencl: fold the gpt-oss MoE bias adds into swiglu_oai
c5ef5073
wanghqc opencl: fold the MoE down-projection bias into the combine
79b151d7
wanghqc wanghqc force pushed to 79b151d7 43 days ago
lhez lhez marked this pull request as ready for review 30 days ago
lhez lhez requested a review 30 days ago
max-krasnyansky
max-krasnyansky approved these changes on 2026-08-21
lhez
lhez approved these changes on 2026-08-21
lhez lhez merged 3af988fa into master 29 days ago

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone