Ollama has introduced preview support for Apple MLX, bringing deep architectural acceleration to open-weight artificial intelligence models running natively on Mac devices.
The update reduces memory footprints and dramatically speeds up token generation for developers running local coding assistants and private terminal agents.