#mlx
- Ollama pre-release runs models on MLX by default on Apple Silicon Ollama tools
- Ollama v0.24.0: Codex App Integration and MLX Sampler Improvements Ollama tools
- Ollama v0.30.10: Cohere Command A and North Models on Apple Silicon via MLX Ollama tools
- Ollama v0.31.1: Gemma 4 Nearly 90% Faster on Apple Silicon via MTP Ollama tools
- Ollama v0.31.2: MLX small-batch matmul kernel, llama.cpp build 9840, CUDA updates tools
- llama.cpp b10603-b10615 — GLM-4.5-Air MTP, Deepseek 4 -sm tensor, Metal flash-attn vec tuning for M1 Pro/M2 Ultra/M5 Max ggml-org tools
- Ollama v0.33.1 adds Qwen3.8 Flash Next support and structured-output mlxrunner Ollama tools
- Ollama v0.34.1 graduates MLX safetensors model creation and speeds up /api/tags tenfold Ollama tools