Ollama Updates MLX Runtime and Supports New Models
Ollama's MLX runtime now runs model architectures on Apple Silicon devices by default, and the platform has added support for several new models.

- Ollama's MLX now supports the models qwen3.8, gemma4, qwen3.6, and qwen3.5.
- According to Ollama (GitHub releases), the company's MLX runtime now runs model architectures on Apple Silicon devices by default.
- The addition of Apple Silicon support expands Ollama's user base and allows developers to integrate the platform's machine learning capabilities into their projects more easily.
What happened
According to Ollama (GitHub releases), the company's MLX runtime now runs model architectures on Apple Silicon devices by default. This update enables developers to deploy Ollama's machine learning models on Apple devices without modification.
Why it matters
The addition of Apple Silicon support expands Ollama's user base and allows developers to integrate the platform's machine learning capabilities into their projects more easily.
Key details
Ollama's MLX now supports the models qwen3.8, gemma4, qwen3.6, and qwen3.5. Additionally, decision models Nimble, tev1, clef, and clef-flash are available on the platform. The company has also added support for the embedding model, embeddinggemma-2. Ollama is continuing to test and enable additional models on its MLX runtime.
Sources
Related stories

Mistral's Le Chonk Model Challenges Top AI Models
Mistral releases a new AI model, Le Chonk, with 1 trillion parameters, optimized for various domains, and available for customization.

Telecom Operators Shift to Open Models for AI
NVIDIA's report reveals a significant 89% shift towards open models among telecom operators, enabling trust, control, and customization of AI capabilities.

OpenAI Releases Mathematical Breakthroughs
The AI research organization publishes a batch of 722 mathematical manuscripts, covering 372 result families and solving hundreds of open questions.

edithero benchmark for long-horizon 3d editing
The EditHero benchmark is introduced to evaluate and improve the reliability of long-horizon, part-level 3D editing.