Apple GPU MLX

imagem 35

Ollama v0.32.5-rc0 Released with MLX Update

The latest release candidate of Ollama, version 0.32.5-rc0, brings an update to the MLX backend. This improvement enhances compatibility and performance for users running Ollama on Apple Silicon hardware. Self-hosted users can expect better integration with the latest MLX features, ensuring a smoother inference experience. As a pre-release, it’s an opportunity to test the updated […]

Ollama v0.32.5-rc0 Released with MLX Update Read More »

imagem 34

Ollama 0.32.4 Brings Laguna to Apple GPUs, Speeds Up Qwen3 MoE, and Refines Speculative Decoding

The latest point release of Ollama, version 0.32.4, expands hardware support by enabling the Laguna model series on Apple GPUs through the MLX engine. This means self-hosted users with Apple Silicon can now run these models with full hardware acceleration, tapping into the performance and efficiency of the Metal-backed MLX runtime that has already proven

Ollama 0.32.4 Brings Laguna to Apple GPUs, Speeds Up Qwen3 MoE, and Refines Speculative Decoding Read More »