What's Changed
Models run on MLX on Apple Silicon by default
In this release, on Apple Silicon devices, model architectures supported by the MLX runtime will automatically run on MLX.
ollama pull qwen3.8
ollama run qwen3.8
During the pre-release we will be testing and enabling additional models.
Full Changelog: v0.34.4...v0.40.0-rc0