b9816
B9816 is a release build number for llama.cpp, an open-source project that enables efficient large language model inference in C/C++ on consumer hardware. This release likely contains bug fixes, perfo
Knowledge catalogue
B9816 is a release build number for llama.cpp, an open-source project that enables efficient large language model inference in C/C++ on consumer hardware. This release likely contains bug fixes, perfo
Build b9784 is a release of llama.cpp, a C/C++ project that enables large language model inference with minimal setup on a wide range of hardware. The release includes pre-compiled binaries for multip
llama.cpp b9556 release adds support for AMD RDNA3.5 graphics hardware (gfx1152 and gfx1153) in its HIP backend . The release includes compiled binaries for multiple platforms including macOS, Linux,
b9523 is a release build of llama.cpp, an open-source tool for LLM inference in C/C++ that enables language model execution with minimal setup on a wide range of hardware locally and in the cloud. The
b9529 is a release tag for llama.cpp, a C/C++ implementation of LLM inference that enables running large language models locally on consumer hardware. This specific release represents a particular ver
b9519 is a release build identifier for llama.cpp, an open-source C/C++ project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware . This specif
B9530 is a release version of llama.cpp, a tool for LLM inference in C/C++ . It allows users to run large language models on everyday consumer hardware without expensive GPUs or cloud infrastructure .
llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Build b9515 is an interme
llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and high performance across diverse hardware . Build b9518 is a release version from the llama.cpp project, which serv
B9471 is a build release of llama.cpp, an open-source C/C++ implementation for LLM inference that enables running large language models on consumer hardware with optimized performance. This intermedia
Ollama 0.30 provides improved compatibility and performance using llama.cpp, augments the MLX engine on Apple Silicon for broader hardware support, and brings support for a wider range of models inclu
B9468 is an intermediate build release of llama.cpp, the open-source C/C++ library that enables large language model inference on consumer hardware. This release continues the project's rapid developm
b9469 is an intermediate build release of llama.cpp, a C/C++ implementation that enables large language model inference on consumer hardware with minimal dependencies. Build releases like b9469 repres
B9403 is an intermediate build release of llama.cpp, an open-source project enabling LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cl
b9410 is a release of llama.cpp, a C/C++ project that enables LLM inference with minimal setup and state-of-the-art performance on various hardware platforms. This build identifier represents a specif
B9222 is a llama.cpp release that adds support for the TRI (Triangle) operation in the Hexagon HTP backend with HVX kernel additions . The release includes optimizations for Hexagon hardware accelerat
B9208 is a build release of llama.cpp, an open-source C/C++ project that enables efficient large language model inference on diverse hardware platforms. The llama.cpp project focuses on optimized LLM
b9221 is an intermediate build release from the llama.cpp project, which is a C/C++ implementation enabling efficient LLM inference on consumer hardware. The release includes platform-specific binarie
B9050 is a release build of llama.cpp, an open-source C/C++ project that enables LLM inference with minimal setup and high performance on diverse hardware platforms. The project provides LLM inference
Release b9061 is a build version of llama.cpp, a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance across various hardware platforms . As a spec
Release b9063 is a build of llama.cpp, a project for LLM inference in C/C++. llama.cpp enables efficient large language model execution on consumer hardware through optimized implementations and quant
B9055 is the latest version of llama.cpp, released on May 7, 2026. Llama.cpp is an LLM inference framework implemented in C/C++ that enables efficient large language model execution with broad hardwar
b9056 is a release of llama.cpp, a C/C++ implementation for LLM inference . The project enables users to run LLaMA models on consumer hardware without expensive GPUs or cloud infrastructure . This rel
b9047 is a build release of llama.cpp, an open-source C/C++ library for running large language model inference on consumer hardware. The release includes pre-compiled binaries for multiple platforms i
This release adds OpenCL GPU acceleration support for the IQ4_NL quantization format in llama.cpp, enabling more efficient inference of quantized language models on compatible hardware. IQ4_NL is a 4-
Release b8934 of llama.cpp was released on April 26, 2026, and includes improvements to hexagon hardware support, specifically guarding HMX clock requests for v75+ platforms. The release provides bina
B8933 is a release from llama.cpp, a project for LLM inference in C/C++. The project aims to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locall
b8931 is a release version of llama.cpp published on April 25, 2026 . llama.cpp is an LLM inference framework in C/C++ that enables running large language models efficiently on various hardware. The r
Based on available information, b9910 is a release tag from the llama.cpp project, an open-source C/C++ implementation for running large language model inference locally on consumer hardware. Llama.cp
b9923 is a release build of llama.cpp, an open-source C/C++ project for LLM inference with minimal setup and state-of-the-art performance on various hardware . The specific b9923 build includes binary
Release b9892 is a version identifier for llama.cpp, an open-source C/C++ framework for running large language model inference on consumer hardware. This specific build (b9892) represents a snapshot o
Release b9873 is a version of llama.cpp, a C/C++ project designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. The
Release b9852 is a build version of llama.cpp, a C/C++ implementation that enables large language model inference with minimal setup and state-of-the-art performance across diverse hardware platforms
llama.cpp is a C/C++ implementation for LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. Release b9844 is an intermediate build of the llama.cpp project f
B9847 is a build release in the llama.cpp project, which is an open-source tool that enables LLM inference with minimal setup on a wide range of hardware . This specific build release likely contains
Build b9829 is an intermediate release of llama.cpp, the open-source C/C++ project that enables users to run large language models on consumer hardware without expensive GPUs or cloud infrastructure.
B9826 is a release of llama.cpp, an LLM inference project written in C/C++ . llama.cpp enables users to run large language models on consumer hardware without expensive GPUs or cloud infrastructure .
llama.cpp is an open-source project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware . Build b9541 is an intermediate release version of llama
B9544 is an intermediate build release of llama.cpp, an open-source C/C++ library for running large language model inference. Llama.cpp uses the GGUF model format and supports multiple hardware backen
Ollama v0.30.4 is a patch release within the v0.30 series, which offers improved compatibility and performance using llama.cpp, augments the MLX engine on Apple Silicon with wider hardware support, an
Release b9464 is a version of llama.cpp, a project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. This specific release likely contains upd
b9431 is a release commit of llama.cpp, an open-source C/C++ implementation for running large language model inference efficiently on consumer hardware. Based on the source material, this release like
B9436 is an intermediate build release of llama.cpp, the open-source C/C++ implementation for running large language model inference on consumer hardware. This build typically includes updates to mode
b9368 is an intermediate build release of llama.cpp, a C/C++ implementation for efficient LLM inference on consumer hardware. The release follows the project's rapid development cycle where multiple b
B9320 is a release of llama.cpp, a C/C++ implementation for enabling LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. The project
B9251 is a build identifier for a release in the llama.cpp project, which enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware - locally and in the clo
b9197 is a build release of llama.cpp , an open-source C/C++ implementation that enables efficient large language model inference on various hardware platforms. The release includes cross-platform bin
b9159 is a release of llama.cpp published on May 14, 2026 . llama.cpp is a C/C++ implementation of large language model inference that enables efficient LLM execution on consumer hardware with minimal
b9172 is a release of llama.cpp that includes binaries for macOS, Linux, Android, Windows, and openEuler platforms with support for various hardware configurations including CPU, Vulkan, CUDA, ROCm, O
Release b9142 is an intermediate build of llama.cpp, a C/C++ library for running large language model inference on consumer hardware. llama.cpp performs inference on various large language models and
b9084 is a release build of llama.cpp, an open-source C/C++ library for LLM inference that enables running large language models on consumer hardware. llama.cpp is a software library that performs inf
llama.cpp is an LLM inference implementation in C/C++ that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud . Build b9079
B9018 is the latest version of llama.cpp released on May 4, 2026. Llama.cpp is a project for LLM inference in C/C++ , providing efficient large language model execution with broad hardware support inc
B9002 is a build release of llama.cpp, a C/C++ implementation for LLM inference. The release includes compiled binaries and artifacts for multiple platforms and hardware configurations, as part of the
llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Release b8995 is a build/versio
b8978 is a release of llama.cpp, a project designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. The llama.cpp project uses sequential build
B8960 is a build release of llama.cpp, an open-source C/C++ project for running large language models locally with minimal setup and optimized performance across various hardware platforms. It is one
B8963 is a build release of llama.cpp, the C/C++ implementation of LLM inference. Llama.cpp is an open-source project that enables efficient language model execution on consumer hardware with minimal
B8924 is a release build number from the llama.cpp project, a C/C++ framework for efficient large language model inference on consumer hardware. The llama.cpp project publishes multiple releases in a
B8902 is a release build of llama.cpp, an open-source C/C++ framework for running large language model inference on consumer hardware with minimal dependencies. As an intermediate build release from t