llama-quant : fail early on missing imatrix, refactor type selection, code cleanup #19770
ddh0
changed the title quantize : refactor llama-quant.cpp quantize : refactor llama-quant.cpp (imatrix fail-early) 179 days ago
ddh0
changed the title quantize : refactor llama-quant.cpp (imatrix fail-early) quantize : fail-early on missing imatrix; refactor + optimize 174 days ago
ddh0
marked this pull request as ready for review 174 days ago
ddh0
marked this pull request as draft 172 days ago
ddh0
changed the title quantize : fail-early on missing imatrix; refactor + optimize quantize : refactor; use quantization work scheduler for faster, more efficient quantization 168 days ago
quantize : imatrix-fail early + code cleanup
decff8b5
ddh0
force pushed
from
e314fa39
to
decff8b5
168 days ago
ddh0
changed the title quantize : refactor; use quantization work scheduler for faster, more efficient quantization quantize : imatrix fail-early, begin code cleanup 168 days ago
ddh0
changed the title quantize : imatrix fail-early, begin code cleanup llama-quant : fail early on missing imatrix, refactor type selection, code cleanup 168 days ago
ddh0
marked this pull request as ready for review 168 days ago
Merge branch 'ggml-org:master' into llama-quant-refactor-2
49fec408
Merge branch 'ggml-org:master' into llama-quant-refactor-2
857d54e1
Merge branch 'ggml-org:master' into llama-quant-refactor-2
8f5c5ea9
Merge branch 'ggml-org:master' into llama-quant-refactor-2
3e888ba2
fix manual override printing
b2b5aa3e
Merge branch 'ggml-org:master' into llama-quant-refactor-2
684ad56c
Merge branch 'ggml-org:master' into llama-quant-refactor-2
e1f9e8c7
revert header changes per @ggerganov
4b7ebed5
remove old #includes
a5e9df2c
clarify naming
ec5b7cb5
Merge branch 'ggml-org:master' into llama-quant-refactor-2
fb38b8cc
fix per barto
f1800552
ggerganov
merged
1dab5f5a
into master 162 days ago
ddh0
deleted the llama-quant-refactor-2 branch 162 days ago
Assignees
No one assigned
Login to write a write a comment.
Login via GitHub