llama.cpp
8e8e2007 - server: add --models-memory-max parameter to allow dynamically unloading models when they exceed a memory size threshold

Commit
113 days ago
server: add --models-memory-max parameter to allow dynamically unloading models when they exceed a memory size threshold
Author
Committer
Parents
Loading