Ollama is supercharged by MLX's unified memory use on Apple Silicon
…Similarly, for decode performance, version 0.18 managed 58 tokens, versus 112 for version 0.19 with MLX. As part of the same update, Ollama has upgraded its cache for efficiency, with…