llama.cpp
e57dc620 - llama: Add support for Gemma2ForCausalLM (#8156)

Commit

346 days ago

llama: Add support for Gemma2ForCausalLM (#8156) * Inference support for Gemma 2 model family * Update convert-hf-to-gguf.py, constants, and tensor mappings * cleanup * format fix * Fix special token vocab bug * Don't add space prefix * fix deleted lines * Update src/llama.cpp Co-authored-by: slaren <slarengh@gmail.com> * Add model type names * Add control vector * Fix model type identification --------- Co-authored-by: Andrei Betlen <abetlen@gmail.com> Co-authored-by: slaren <slarengh@gmail.com>

References

#8156 - Add support for Gemma2ForCausalLM

Author

pculliton

Parents

a27aa50a

Files4

convert-hf-to-gguf.py
gguf-py/gguf
- constants.py
- tensor_mapping.py
src
- llama.cpp

llama.cpp e57dc620 - llama: Add support for Gemma2ForCausalLM (#8156)

llama.cpp
e57dc620 - llama: Add support for Gemma2ForCausalLM (#8156)