Laguna model

imagem 40

Ollama v0.32.5 Fixes NVFP4 Output Quality Bug on Apple Silicon

Ollama v0.32.5 Release Notes The latest point release, Ollama v0.32.5, resolves a significant bug affecting users running NVFP4‑based models such as Laguna on Apple Silicon Macs. A flaw in the MLX Metal backend could previously degrade output quality, leading to less accurate or lower‑fidelity responses from these models. With the fix in place, self‑hosted users […]

Ollama v0.32.5 Fixes NVFP4 Output Quality Bug on Apple Silicon Read More »

imagem 34

Ollama 0.32.4 Brings Laguna to Apple GPUs, Speeds Up Qwen3 MoE, and Refines Speculative Decoding

The latest point release of Ollama, version 0.32.4, expands hardware support by enabling the Laguna model series on Apple GPUs through the MLX engine. This means self-hosted users with Apple Silicon can now run these models with full hardware acceleration, tapping into the performance and efficiency of the Metal-backed MLX runtime that has already proven

Ollama 0.32.4 Brings Laguna to Apple GPUs, Speeds Up Qwen3 MoE, and Refines Speculative Decoding Read More »