llama.cpp
Only use Q6_K for output weights if tensor size is multiple of 256
#1932
Merged

Commits
  • Only use Q6_K for output weights if tensor size is multiple of 256
    Iwan Kawrakow committed 3 years ago
  • Fixed copy/paste mistake
    Iwan Kawrakow committed 3 years ago
Loading