mirror of
https://github.com/ggml-org/llama.cpp.git
synced 2026-10-01 02:17:42 -05:00
The manager could narrow by provider, context and modality only. It now also asks for tool use or for reasoning, which the Hub chat template answers for a local model, and for models that have a draft sidecar to speculate with. The provider menu says how many repos each provider contributes and steps aside when there is only one to choose between. The Hub details cache becomes a SvelteMap so a filter re-runs when a template arrives after the row mounts, and the discover entry's effect consumes its flag untracked instead of invalidating itself. Assisted-by: pi:llama.cpp/DeepSeek-V4.1-Flash