llama.cpp
0e797c2f - llm : support Adept Persimmon 8B (#3410)

Commit

2 years ago

llm : support Adept Persimmon 8B (#3410) * Produces garbage output * wip: correct tensors up to RoPE * correct tensors thru RoPE * Correct outputs through masked & softmax'd KQ * fp32 works * Rename adept->persimmon * Produces correct outputs * clean up convert scripts * remove printing logic from ggml.c * remove prints from llama.cpp & fix merge * trivial cleanups * Add offload funcs * update conversion script to directly take adept artifacts rather than .saftensors file * Fix norm eps bug * Support sqr and concat on metal, persimmon-8b-q4 runs correctly * Small changes from review * Formatting changes * Minor changes to conversion script * Remove old script * Fix editorconfig formatting * Fix build * add overlooked offload code ggml-ci

References

#3410 - Support Adept Persimmon 8b

Author

phillip-kravtsov

Parents

3a716b4d

llama.cpp 0e797c2f - llm : support Adept Persimmon 8B (#3410)

llama.cpp
0e797c2f - llm : support Adept Persimmon 8B (#3410)