mirror of
https://github.com/ollama/ollama.git
synced 2026-09-11 15:27:46 +00:00
With the new version of GGML in #12245, KV cache quantization no longer causes a fallback to CPU. |
||
|---|---|---|
| .. | ||
| ggml | ||
| gguf | ||
| util/bufioutil | ||
| config.go | ||