Gemma 4 is now significantly faster in Ollama 0.31 on Apple Silicon via multi-token prediction (MTP), powered by MLX. Performance is up to 90% faster when used with coding agents, as measured on the Aider polyglot benchmark.
Need help?
Contact usGemma 4 is now significantly faster in Ollama 0.31 on Apple Silicon via multi-token prediction (MTP), powered by MLX. Performance is up to 90% faster when used with coding agents, as measured on the Aider polyglot benchmark.