llama.cpp
8e8e2007
- server: add --models-memory-max parameter to allow dynamically unloading models when they exceed a memory size threshold
Go
Login via GitHub
Home
Pricing
FAQ
Install
Login
via GitHub
Commit
View On
GitHub
Commit
113 days ago
server: add --models-memory-max parameter to allow dynamically unloading models when they exceed a memory size threshold
References
#22284 - server: router fix model unload reload deadlock
Author
0cc4m
Committer
0cc4m
Parents
82209efb
Loading