

Ollama v0.19
Massive local model speedup on Apple Silicon with MLX
What is Ollama v0.19?
Ollama v0.19 rebuilds Apple Silicon inference on top of MLX, bringing much faster local performance for coding and agent workflows. It also adds NVFP4 support and smarter cache reuse, snapshots, and eviction for more responsive sessions.
From the website
Ollama is now powered by MLX on Apple Silicon in preview · Ollama Blog
Today, we're previewing the fastest way to run Ollama on Apple silicon, powered by MLX, Apple's machine learning framework.
Visit Website →Screenshots
?
No comments yet. Be the first!




