mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-08-02 08:00:48 -05:00
* server: document --n-predict * server: ensure client request cannot override n_predict if set * server: fix print usage LF in new --n-predict option
120 KiB
120 KiB