

Ollama v0.19
Massive local model speedup on Apple Silicon with MLX
什么是 Ollama v0.19?
Ollama v0.19 rebuilds Apple Silicon inference on top of MLX, bringing much faster local performance for coding and agent workflows. It also adds NVFP4 support and smarter cache reuse, snapshots, and eviction for more responsive sessions.
截图
?
还没有评论,来抢沙发吧!
X 上关于 Ollama v0.19 的真实讨论
去 X 发帖



