llama.cpp
9ae4143b - model : add dots.llm1 architecture support (#14044) (#14118)

Commit

187 days ago

model : add dots.llm1 architecture support (#14044) (#14118) Adds: * Dots1Model to convert_hf_to_gguf.py * Computation graph code to llama-model.cpp * Chat template to llama-chat.cpp to detect this model's template. --- The model is called "dots.llm1" (I decided to shorten it to dots1 or DOTS1 in the code generally) architecture. The only models that exist as of writing of this commit that follow this architecture are "dots.llm1.inst" and "dots.llm1.base" from here: * https://huggingface.co/rednote-hilab/dots.llm1.inst * https://huggingface.co/rednote-hilab/dots.llm1.base The model architecture is a combination of Qwen and Deepseek parts, as seen here: https://github.com/huggingface/transformers/blob/ffe12627b4e84489d2ab91dd0ec00614855edc79/src/transformers/models/dots1/modular_dots1.py

References

#14118 - llama-model : add dots.llm1 architecture support (#14044)

Author

Noeda

Parents

c311ac66

llama.cpp 9ae4143b - model : add dots.llm1 architecture support (#14044) (#14118)

llama.cpp
9ae4143b - model : add dots.llm1 architecture support (#14044) (#14118)