llama.cpp
732707df
- quantize: cap working memory size to avoid loading big tensors onto RAM (#27795)
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
6 days ago
quantize: cap working memory size to avoid loading big tensors onto RAM (#27795)
References
#27795 - quantize: cap working memory size to avoid loading big tensors onto RAM
Author
ngxson
Parents
cb300598
Loading