mirror of
https://github.com/ollama/ollama.git
synced 2026-08-04 14:56:15 +00:00
Update llama.cpp to pick up upstream Laguna implementation and remove Ollama's local Laguna implementation. Retain a narrow Metal-only scaling workaround for routed-MoE prompt overflow. Translate older Ollama GGUF attention-gate and SWA metadata names so existing models continue to load. |
||
|---|---|---|
| .. | ||
| parsers | ||
| renderers | ||