mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-07-25 20:21:03 -05:00
* server : clear checkpoints upon prompt clear * server : move the prompt state data to the server_prompt_cache Assisted-by: pi:llama.cpp/Qwen3.6-27B * server : handle batched slot being cleared
63 KiB
63 KiB