Switch language한국어
Back to the list

Running local models on Macs gets faster with Ollama's MLX support

TL;DR AI

Key summary

1 min read
  1. Ollama added support for Apple’s open source MLX framework to run local models.

  2. Ollama also improved caching and added Nvidia NVFP4 model compression support.

  3. The new features, in preview in Ollama 0.19, currently support only Qwen3.5 35B and target M1-or-later Macs with at least 32GB RAM.

Read the original