llama.cpp Release b10159 Brings FWHT Kernel to Metal Backend
We’re excited to announce a new build of llama.cpp (b10159), which includes a performance-boosting addition for Apple Silicon users. This release introduces the Fast Walsh-Hadamard Transform (FWHT) kernel into the Metal backend, improving the efficiency of certain mathematical operations during model inference. While the change is low-level, it helps streamline computation on macOS and iOS […]
llama.cpp Release b10159 Brings FWHT Kernel to Metal Backend Read More »

