llama.cpp
server : add `n_discard` parameter to specify the number of tokens to discard when context is shifted
#6300
Merged

Commits
  • server : add `n_discard` parameter to specify the number of tokens to discard when context is shifted
    kaetemi committed 2 years ago
Loading