llama.cpp Build b10227 Brings Qwen3 Parser and Improved Tool Calling Support

imagem 3

The latest release of llama.cpp, build b10227, introduces a specialized parser for the Qwen3 model family, enhancing how the chat interface handles tagged thinking tools. This update refactors internal chat processing and adds a permute helper, streamlining the codebase for future expansions. Self-hosted users will benefit from more accurate tool calling, as the update adds support for omitting certain tool call tags and adjusts trigger patterns to ensure functions like <function are properly recognized. These changes, alongside updated tool delimiters and a note for the Qwen3-Coder variant, collectively improve the reliability of interactive chat sessions and automated function calls when running Qwen3 models locally.

In addition to the parser improvements, the release provides downloadable binaries for a wide range of platforms, ensuring that users can deploy llama.cpp on their preferred operating system and hardware. Below is the complete list of available builds:

Website:

macOS/iOS:

Linux:

Android:

Windows:

openEuler:

  • Currently disabled — openEuler builds are not available at this time
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

UI:

Leave a Comment

Your email address will not be published. Required fields are marked *