#apple-silicon
- PrismML Releases Bonsai 27B: First 27B-Class Model to Run on iPhone PrismML models-llm
- llama.cpp v0.3.0 — first tagged stable release of the b10621 nightly; dots3-note multimodal, GLM-4.5-Air MTP, DeepSeek 4 tensor-split, ggml v0.22.0 ggml-org tools
- Ollama v0.23.1: Gemma 4 MTP Speculative Decoding Delivers 2× Speed on Apple Silicon tools
- Ollama v0.24.0: Codex App Integration and MLX Sampler Improvements Ollama tools
- Ollama v0.30.10: Cohere Command A and North Models on Apple Silicon via MLX Ollama tools
- Ollama v0.31.1: Gemma 4 Nearly 90% Faster on Apple Silicon via MTP Ollama tools
- Ollama v0.31.2: MLX small-batch matmul kernel, llama.cpp build 9840, CUDA updates tools
- llama.cpp b10603-b10615 — GLM-4.5-Air MTP, Deepseek 4 -sm tensor, Metal flash-attn vec tuning for M1 Pro/M2 Ultra/M5 Max ggml-org tools
- Ollama v0.23.3: MLX Runner Fixes and macOS 26 Metal Compatibility Ollama tools