
Open Source
Faster Gemma 4 on MLX with multi-token prediction
Faster Gemma 4 on MLX with multi-token prediction Gemma 4 is now significantly faster in Ollama 0.31. On Apple Silicon, it generates tokens nearly 90% faster on average across a coding-agent benchmark.