

Massive local model speedup on Apple Silicon with MLX
Ollama v0.19 rebuilds Apple Silicon inference on top of MLX, bringing much faster local performance for coding and agent workflows. It also adds NVFP4 support and smarter cache reuse, snapshots, and eviction for more responsive sessions.
아직 댓글이 없어요. 가장 먼저 남겨보세요!
Ollama v0.19에 대한 X의 실제 대화
X에 게시