mirror of
https://github.com/ollama/ollama.git
synced 2026-07-23 09:10:53 -05:00
model: align Laguna with upstream llama.cpp (#17335)
Update llama.cpp to pick up upstream Laguna implementation and remove Ollama's local Laguna implementation. Retain a narrow Metal-only scaling workaround for routed-MoE prompt overflow. Translate older Ollama GGUF attention-gate and SWA metadata names so existing models continue to load.
This commit is contained in:
@@ -1 +1 @@
|
||||
b10069
|
||||
b10091
|
||||
|
||||
Reference in New Issue
Block a user