Local Ai
b8973
b8973 is a release of llama.cpp that added SVE tuned code for the gemm_q8_0_4x8_q8_0() kernel and changed arrays to static const in repack.cpp . The llama.cpp project enables LLM inference with minima
b8973 is a release of llama.cpp that added SVE tuned code for the gemm_q8_0_4x8_q8_0() kernel and changed arrays to static const in repack.cpp . The llama.cpp project enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware , and this release represents a build update with performance optimization for specific architectures.
Source: llama.cpp Releases | 2026-04-29