AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “llama-cpp-releases”

GridTimelineEvolution
658 results
7 May 2026

b9061

Local AiDGX agent

Release b9061 is a build version of llama.cpp, a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance across various hardware platforms . As a spec

b9063

Local AiDGX agent

Release b9063 is a build of llama.cpp, a project for LLM inference in C/C++. llama.cpp enables efficient large language model execution on consumer hardware through optimized implementations and quant

b9066

Local AiDGX agent

llama.cpp b9066 is a release of a C/C++ implementation for large language model (LLM) inference . Based on the release repository structure, b9066 represents a specific build or commit version of the

6 May 2026

b9041


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai
DGX agent

llama.cpp b9041 is a release build of llama.cpp, an LLM inference library written in C/C++ . The build identifier indicates this is an intermediate development release from the ggml-org/llama.cpp proj

b9047

Local AiDGX agent

b9047 is a build release of llama.cpp, an open-source C/C++ library for running large language model inference on consumer hardware. The release includes pre-compiled binaries for multiple platforms i

b9048

Local AiDGX agent

B9048 is a build/release of llama.cpp, an open-source C/C++ implementation for running large language model inference. The release includes optimizations and improvements to the core llama.cpp softwar

5 May 2026

b9028

Local AiDGX agent

B9028 is a release build from the llama.cpp project, which provides LLM inference in C/C++. The llama.cpp project uses sequential build numbers (like b9028) to track intermediate releases of its LLM i

b9030

Local AiDGX agent

B9030 is a release build of llama.cpp, a C/C++ implementation for LLM inference . This build release would contain bug fixes, performance improvements, and new features added to the llama.cpp project

b9031

Local AiDGX agent

Release b9031 of llama.cpp optimizes backend loading by only loading backends when required, as specified in pull request #22290 . The release includes binary builds for multiple platforms including m

b9033

Local AiDGX agent

Release b9033 is a build of llama.cpp with a sync of ggml , published on May 5, 2026. The release provides compiled binaries across multiple platforms including macOS, Linux, Android, Windows, and ope

4 May 2026

b9014

Local AiDGX agent

llama.cpp enables LLM inference in C/C++ , and release b9014 represents a build revision or commit snapshot from the llama.cpp repository on GitHub. This release would include bug fixes, feature impro

b9015

Local AiDGX agent

The search results show recent llama.cpp releases but not the specific b9015 release details. Based on the pattern found (b9010 is mentioned as the latest release from the search results), b9015 is li

b9016

Local AiDGX agent

Release b9016 is a build of llama.cpp, an LLM inference implementation in C/C++ . This build designation represents a specific intermediate version in the project's development cycle, similar to other

b9018

Local AiDGX agent

B9018 is the latest version of llama.cpp released on May 4, 2026. Llama.cpp is a project for LLM inference in C/C++ , providing efficient large language model execution with broad hardware support inc

b9019

Local AiDGX agent

b9019 is a release of llama.cpp, a C/C++ implementation for LLM inference. As a release tag from the ggml-org/llama.cpp repository, it represents a specific commit or version of the project that inclu

b9020

Local AiDGX agent

b9020 is a build release of llama.cpp (a C/C++ implementation of LLM inference) released on May 4, 2026, with precompiled binaries available for multiple platforms including macOS, Android, OpenEuler,

b9022

Local AiDGX agent

B9022 is a release of llama.cpp created on May 4, 2026, with commit d8794ee signed with GitHub's verified signature. The llama.cpp project is the main playground for developing new features for the gg

b9025

Local AiDGX agent

b9025 is a release from llama.cpp, a C/C++ project for LLM inference . The release uses a build numbering system for intermediate versions of the library. As a llama.cpp release, it likely includes up

3 May 2026

b9012

Local AiDGX agent

b9012 is a build version from the llama.cpp GitHub repository, which is an LLM inference implementation in C/C++ . The release likely contains updates, bug fixes, and improvements to the llama.cpp inf

2 May 2026

b9000

Local AiDGX agent

b9000 is a build release of llama.cpp, a C/C++ implementation for LLM inference . The release follows the project's rapid development cycle where multiple releases can be published in a single day . A

b9002

Local AiDGX agent

B9002 is a build release of llama.cpp, a C/C++ implementation for LLM inference. The release includes compiled binaries and artifacts for multiple platforms and hardware configurations, as part of the

b9004

Local AiDGX agent

llama.cpp is an LLM inference framework in C/C++ that enables efficient local execution of large language models. Build b9004 is a recent release with optimizations and improvements for supporting var

b9008

Local AiDGX agent

B9008 is a build release of llama.cpp from May 2, 2026. Llama.cpp is a C/C++ implementation of LLM inference designed to enable large language model inference with minimal setup and high performance a

b9009

Local AiDGX agent

Build b9009 is an incremental release of llama.cpp from May 2026, continuing the rapid development cycle that characterized April 2026's updates with tensor parallelism, 1-bit quantization, and expand

b9010

Local AiDGX agent

b9010 is a build release of llama.cpp, an open-source project for LLM inference in C/C++ . As a numbered build tag in the llama.cpp release system, it represents a specific development build containin

1 May 2026

b8995

Local AiDGX agent

llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Release b8995 is a build/versio

b8999

Local AiDGX agent

llama.cpp b8999 is a build release from the llama.cpp project, which provides LLM inference in C/C++ . This build represents an intermediate development version in the project's rapid release cycle. T

30 Apr 2026

b8982

Local AiDGX agent

The search did not return specific details about release b8982. Based on the llama.cpp project context from the search results, b8982 is a build/release version from the llama.cpp repository that foll

b8983

Local AiDGX agent

b8983 is a release of llama.cpp, a C/C++ implementation for LLM inference. The release represents a specific build/commit version in the active development of the llama.cpp project, which supports run

b8984

Local AiDGX agent

B8984 is a build/release version of llama.cpp, an open-source C/C++ library for large language model inference. Llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a

b8987

Local AiDGX agent

Release b8987 of llama.cpp updates cpp-httplib to version 0.43.2 . This is a minor maintenance release from the llama.cpp project, which provides C/C++ inference for large language models. The release

b8989

Local AiDGX agent

Based on the available search results, I cannot locate specific details about the b8989 release. However, b8989 is a commit hash from the llama.cpp repository releases. llama.cpp is a C/C++ implementa

b8990

Local AiDGX agent

llama.cpp b8990 is a release version of a C/C++ LLM inference tool from the ggml-org project. The release likely contains bug fixes, performance improvements, or new features for local language model

b8991

Local AiDGX agent

llama.cpp is a C/C++ implementation for LLM inference , and b8991 is a specific release version from the project's ongoing development cycle. The search results indicate that the project doesn't follo

b8992

Local AiDGX agent

b8992 is a build release from the llama.cpp project, which publishes multiple releases in a single day . llama.cpp enables LLM inference with minimal setup and state-of-the-art performance on a wide r

29 Apr 2026

b8969

Local AiDGX agent

b8969 is a release build of llama.cpp (commit bdc9c74), published on April 29, 2026. Llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art perfor

b8970

Local AiDGX agent

b8970 is a release of llama.cpp, a C/C++ project for large language model inference. The naming convention using a build hash (b8970) is typical of llama.cpp's continuous release pattern rather than t

b8971

Local AiDGX agent

b8971 is a build release of llama.cpp, an open-source C/C++ project for LLM inference . As a numbered build identifier from the llama.cpp repository, this release likely contains bug fixes, performanc

b8972

Local AiDGX agent

b8972 is a release of llama.cpp, a C/C++ implementation for LLM inference . The release identifier follows llama.cpp's development versioning scheme with frequent builds published to the GitHub releas

b8973

Local AiDGX agent

b8973 is a release of llama.cpp that added SVE tuned code for the gemm_q8_0_4x8_q8_0() kernel and changed arrays to static const in repack.cpp . The llama.cpp project enables LLM inference with minima

b8978

Local AiDGX agent

b8978 is a release of llama.cpp, a project designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. The llama.cpp project uses sequential build

b8979

Local AiDGX agent

B8979 is a llama.cpp release that added SVE (Scalable Vector Extension) tuned code for the gemm_q8_0_4x8_q8_0() kernel and made static const changes to repack.cpp . The release includes pre-built bina

b8981

Local AiDGX agent

b8981 is a release version of llama.cpp, a C/C++ implementation for LLM inference . Build versions in llama.cpp releases typically contain updates to the underlying ggml library, performance optimizat

28 Apr 2026

b8953

Local AiDGX agent

Release b8953 of llama.cpp adds Q1_0 quantization support for WebGPU, including fast matmul and matvec kernels and optimized shared memory initialization. The release was published on April 28 and inc

b8954

Local AiDGX agent

llama.cpp is a C/C++ implementation for LLM inference. Build b8954 is a release version from the ggml-org/llama.cpp repository that likely includes bug fixes, optimizations, or new features added sinc

b8957

Local AiDGX agent

b8957 is a release of llama.cpp that includes spec parameter refactoring changes , featuring modifications to parameter handling and model specifications. The llama.cpp project enables LLM inference w

b8958

Local AiDGX agent

b8958 is a release of llama.cpp, a C/C++ library for LLM inference . The project uses rapid release cycles with frequent version tags tracking incremental updates and improvements to the codebase. Thi

b8960

Local AiDGX agent

B8960 is a build release of llama.cpp, an open-source C/C++ project for running large language models locally with minimal setup and optimized performance across various hardware platforms. It is one

b8963

Local AiDGX agent

B8963 is a build release of llama.cpp, the C/C++ implementation of LLM inference. Llama.cpp is an open-source project that enables efficient language model execution on consumer hardware with minimal

b8964

Local AiDGX agent

b8964 is a release of llama.cpp, a C/C++ library for LLM inference . The release likely includes bug fixes, performance optimizations, or new features as part of the project's continuous development c

b8966

Local AiDGX agent

b8966 is one of the latest releases in the llama.cpp project , a C/C++ implementation for efficient LLM inference. This release represents an intermediate development build in the llama.cpp repository

b8967

Local AiDGX agent

llama.cpp b8967 is an intermediate build release of llama.cpp, a C/C++ implementation for LLM inference with minimal setup and state-of-the-art performance. The project follows a rapid release cycle w

27 Apr 2026

b8943

Local AiDGX agent

Release b8943 of llama.cpp fixes missing exports in llama-common and refactors the common/debug module, moving abort_on_nan from a template parameter to a member of base_callback_data. The release als

b8944

Local AiDGX agent

Release b8944 was published on April 27, 2026 , representing a recent build of the llama.cpp project. Llama.cpp provides LLM inference in C/C++ for running large language models efficiently on consume

b8946

Local AiDGX agent

Release b8946 includes a fix to remove duplicate wo_s scale after build_attn for Qwen3 and LLaMA models , along with other improvements to the llama.cpp library. This is an intermediate build release

b8948

Local AiDGX agent

The search results do not contain specific information about release b8948 itself. Based on the context from llama.cpp releases, b8948 is one of the continuously updated build releases from the llama.

b8950

Local AiDGX agent

The search results don't contain specific details about the b8950 release. Based on the naming convention and context, b8950 is an intermediate build release of llama.cpp, a C/C++ implementation of LL

b8952

Local AiDGX agent

The search results don't contain specific details about release b8952. Based on the information available, I can provide this knowledge base entry: b8952 is a release in the llama.cpp project, an open

26 Apr 2026

b8934

Local AiDGX agent

Release b8934 of llama.cpp was released on April 26, 2026, and includes improvements to hexagon hardware support, specifically guarding HMX clock requests for v75+ platforms. The release provides bina

b8935: opencl: add iq4_nl support (#22272)

Local AiDGX agent

This release adds OpenCL GPU acceleration support for the IQ4_NL quantization format in llama.cpp, enabling more efficient inference of quantized language models on compatible hardware. IQ4_NL is a 4-

← Previous
1…7891011
Next →