llama.cpp
ggml-cpu: use AVX-512 VNNI+VBMI in the Q4_K 8x8 repack GEMM
#29397
Open

ggml-cpu: use AVX-512 VNNI+VBMI in the Q4_K 8x8 repack GEMM #29397

JeremiahM37
JeremiahM37 ggml-cpu: use AVX-512 VNNI+VBMI in the Q4_K 8x8 repack GEMM
51073eff
JeremiahM37 JeremiahM37 requested a review from ggerganov ggerganov 1 day ago
ggml-gh-bot
ggml-gh-bot ggml-gh-bot added draft
github-actions github-actions added ggml
github-actions github-actions marked this pull request as draft 1 day ago
github-actions github-actions removed draft

Login to write a write a comment.

Login via GitHub

Reviewers
Assignees
No one assigned
Labels
Milestone