ollama

imagem 38

Ollama v0.34.0 Brings ChatGPT Desktop Integration and Performance Boosts

The latest Ollama release, v0.34.0, introduces direct integration with ChatGPT Desktop, letting you use your locally hosted Ollama models without leaving the ChatGPT interface. Once set up through the Ollama app on macOS, you can keep your existing workflow while running open models side by side. This update also improves structured output performance on Apple […]

Ollama v0.34.0 Brings ChatGPT Desktop Integration and Performance Boosts Read More »

imagem 15

Ollama 0.34.0-rc1 Adds ChatGPT Desktop Integration and Apple Silicon Optimizations

You can now run Ollama models directly inside ChatGPT Desktop, letting you keep your current workflow while using open models. On macOS, setup is handled right from the Ollama app. The update also brings faster structured output on Apple Silicon, support for tool search and response compaction in OpenAI-compatible clients, and correct image handling through

Ollama 0.34.0-rc1 Adds ChatGPT Desktop Integration and Apple Silicon Optimizations Read More »

imagem 67

Ollama v0.33.2 Restores Dark Mode and Fixes macOS Instance Handling

Ollama v0.33.2 brings back the app’s ability to follow the system appearance, restoring dark mode support for users who depend on it. The macOS application now correctly hands off to an already-running instance instead of starting a second one, preventing duplicate processes and unnecessary resource usage. Additionally, the Claude Desktop proxy has been fixed so

Ollama v0.33.2 Restores Dark Mode and Fixes macOS Instance Handling Read More »

imagem 50

Ollama v0.33.0-rc2 Adds Claude Desktop Integration and Cache Reliability Fixes

Ollama’s latest release candidate, v0.33.0-rc2, brings a handful of meaningful updates for self-hosted users, including tighter Claude Desktop integration and several under-the-hood improvements that make model caching more dependable. The Claude Desktop app now works directly with Ollama. You can toggle individual Ollama models on or off right from the menu bar, and pick any

Ollama v0.33.0-rc2 Adds Claude Desktop Integration and Cache Reliability Fixes Read More »

imagem 30

Ollama v0.32.14-rc0 Adds WebP Transcoding and Qwen Compatibility Fix

The latest release candidate for Ollama, v0.32.14-rc0, includes two focused improvements for self-hosted users. First, llama-server now automatically transcodes WebP images before sending them to the model. This means you can use WebP-format images in your prompts without manually converting them, even if the underlying model does not natively support that format. Second, the Qwen

Ollama v0.32.14-rc0 Adds WebP Transcoding and Qwen Compatibility Fix Read More »

imagem 23

Ollama v0.32.6-rc0 Brings Faster Qwen3.5 and Better OpenAI Streaming

Performance Boost for Qwen3.5 on Apple GPUs Ollama’s latest release candidate speeds up Qwen3.5 on Apple hardware by automatically enabling speculative decoding through the MLX engine, which now leverages the model’s MTP head. The MLX and llama.cpp engines have also been updated to their latest versions, ensuring broader compatibility and performance improvements. Smoothed API Compatibility

Ollama v0.32.6-rc0 Brings Faster Qwen3.5 and Better OpenAI Streaming Read More »

imagem 12

How to Run Ollama AI Models Locally with Podman on Fedora

Running large language models (LLMs) locally has gained traction for development, privacy, and offline testing. Ollama streamlines this process, letting you run models like Llama 3 or Mistral right on your machine. By using Podman on Fedora Linux, you can isolate Ollama within a container. This keeps your system tidy and makes it simple to

How to Run Ollama AI Models Locally with Podman on Fedora Read More »

imagem 40

Ollama v0.32.5 Fixes NVFP4 Output Quality Bug on Apple Silicon

Ollama v0.32.5 Release Notes The latest point release, Ollama v0.32.5, resolves a significant bug affecting users running NVFP4‑based models such as Laguna on Apple Silicon Macs. A flaw in the MLX Metal backend could previously degrade output quality, leading to less accurate or lower‑fidelity responses from these models. With the fix in place, self‑hosted users

Ollama v0.32.5 Fixes NVFP4 Output Quality Bug on Apple Silicon Read More »

imagem 35

Ollama v0.32.5-rc0 Released with MLX Update

The latest release candidate of Ollama, version 0.32.5-rc0, brings an update to the MLX backend. This improvement enhances compatibility and performance for users running Ollama on Apple Silicon hardware. Self-hosted users can expect better integration with the latest MLX features, ensuring a smoother inference experience. As a pre-release, it’s an opportunity to test the updated

Ollama v0.32.5-rc0 Released with MLX Update Read More »

imagem 34

Ollama 0.32.4 Brings Laguna to Apple GPUs, Speeds Up Qwen3 MoE, and Refines Speculative Decoding

The latest point release of Ollama, version 0.32.4, expands hardware support by enabling the Laguna model series on Apple GPUs through the MLX engine. This means self-hosted users with Apple Silicon can now run these models with full hardware acceleration, tapping into the performance and efficiency of the Metal-backed MLX runtime that has already proven

Ollama 0.32.4 Brings Laguna to Apple GPUs, Speeds Up Qwen3 MoE, and Refines Speculative Decoding Read More »