Files
llama.cpp/include
Oliver Simons 748eca633a Resolve -1 to 1024 instead of ctx-len for samplers
Because of backend-sampling we initialize samplers before the complete
llama_context is there. Therefore, we cannot infer the resolved context
length yet at the time we construct the samplers.
2026-08-04 09:02:53 +03:00
..