llama.cpp
e6ec21e6 - ggml-cpu: add always_inline to tinyBLAS_PPC accumulator saves (#20791)

Commit
4 days ago
ggml-cpu: add always_inline to tinyBLAS_PPC accumulator saves (#20791) Explicitly mark save_acc and add_save_Acc with always_inline in tinyBLAS_PPC. This ensures the compiler keeps MMA accumulator disassembly within kernel's register context, preventing un-necessary stask spills. Signed-off-by: Shalini Salomi Bodapati <Shalini.Salomi.Bodapati@ibm.com>
Author
Parents
Loading