llama.cpp Release b10159 Brings FWHT Kernel to Metal Backend

imagem 41

We’re excited to announce a new build of llama.cpp (b10159), which includes a performance-boosting addition for Apple Silicon users. This release introduces the Fast Walsh-Hadamard Transform (FWHT) kernel into the Metal backend, improving the efficiency of certain mathematical operations during model inference. While the change is low-level, it helps streamline computation on macOS and iOS devices with Metal support, contributing to faster and smoother local AI experiences.

The following pre-built packages are available for download:

Website

macOS/iOS

Linux

Android

Windows

openEuler

  • [DISABLED]
  • openEuler x86 (310p)
  • openEuler x86 (910b, ACL Graph)
  • openEuler aarch64 (310p)
  • openEuler aarch64 (910b, ACL Graph)

UI

Leave a Comment

Your email address will not be published. Required fields are marked *