AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
All
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “llama-cpp-releases”

GridTimelineEvolution
658 results
Local Ai

b9581

DGX agent

llama.cpp b9581 is a release that includes optimization for Vulkan backend memory usage, specifically reducing iq1 shared memory usage for mul_mm operations. Released on June 9, 2026 , this build prov

local-aillama-cpp-releases
9 Jun 2026
Local Ai

b9584

DGX agent

The search results don't contain specific information about release b9584. Based on the context from other llama.cpp releases in the results, b9584 is likely an intermediate build version of llama.cpp

local-aillama-cpp-releases
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
9 Jun 2026
Local Ai

b9585

DGX agent

Release b9585 of llama.cpp fixes granite speech model inference by applying embedding scale when deepstack is not used . The release was published on June 9, 2026, and represents a bug fix within the

local-aillama-cpp-releases
9 Jun 2026
Local Ai

b9554: [SYCL] Update compute runtime version to 26.x in docker (#24070)

DGX agent

This release updates the compute runtime version to 26.x within a Docker configuration for SYCL (Syclon Compute Language) support in llama.cpp. The change, addressed in pull request #24070, modernizes

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9555

DGX agent

llama.cpp is an LLM inference tool in C/C++ , and b9555 represents a specific build or version release from the ggml-org/llama.cpp GitHub repository. The project does not follow traditional release pr

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9556

DGX agent

llama.cpp b9556 release adds support for AMD RDNA3.5 graphics hardware (gfx1152 and gfx1153) in its HIP backend . The release includes compiled binaries for multiple platforms including macOS, Linux,

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9558

DGX agent

llama.cpp release b9558 includes a Vulkan optimization that uses cm2 decode_vector for mul_mat_id B matrix loads, allowing vec4 loads and increasing BK to 64, resulting in performance speedups. The re

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9561

DGX agent

B9561 is an intermediate build release of llama.cpp, the C/C++ implementation of large language model inference. Llama.cpp releases use build identifiers (b-numbers) to track development versions betw

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9563

DGX agent

b9563 is a release build of llama.cpp, an open-source software library for large language model inference that is co-developed alongside the GGML tensor library. This intermediate build includes vario

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9564

DGX agent

The search results did not provide specific information about the b9564 release. Based on the context of llama.cpp releases and the pattern observed with nearby releases (b9542, b9543, b9544, etc.), I

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9568

DGX agent

Based on available search results, I cannot find specific details about the b9568 release. B9568 is a build number from llama.cpp, an LLM inference implementation in C/C++ . This release likely contai

local-aillama-cpp-releases
8 Jun 2026
Local Ai

b9547

DGX agent

The search did not return specific details about the b9547 release. Based on the repository context and release naming convention, b9547 is a build number for llama.cpp, an open-source C/C++ implement

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9548

DGX agent

Release b9548 of llama.cpp includes a spec fix for vocabulary compatibility checking , addressing issues related to model format validation. This build provides compiled binaries for multiple platform

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9549

DGX agent

Release b9549 of llama.cpp adds support for the Gemma4 MTP model architecture . The release includes pre-built binaries for multiple platforms including macOS, Linux, Android, and Windows with various

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9550

DGX agent

b9550 is a release of llama.cpp, a C/C++ implementation for LLM inference . The release represents an intermediate build in the llama.cpp development cycle, which releases frequently without tradition

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9551

DGX agent

llama.cpp is a project for LLM inference in C/C++ . Release b9551 is a build version from the llama.cpp project's GitHub releases, following the project's rapid development cycle where multiple releas

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9538

DGX agent

b9538 is a release version of llama.cpp, an open-source C/C++ project for LLM inference. As a commit hash-based release in the llama.cpp repository, it represents a specific point in the project's dev

local-aillama-cpp-releases
6 Jun 2026
Local Ai

b9541

DGX agent

llama.cpp is an open-source project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware . Build b9541 is an intermediate release version of llama

local-aillama-cpp-releases
6 Jun 2026
Local Ai

b9542

DGX agent

B9542 is a build/release version of llama.cpp, the C/C++ implementation of large language model inference designed to enable LLM inference with minimal setup and state-of-the-art performance on variou

local-aillama-cpp-releases
6 Jun 2026
Local Ai

b9544

DGX agent

B9544 is an intermediate build release of llama.cpp, an open-source C/C++ library for running large language model inference. Llama.cpp uses the GGUF model format and supports multiple hardware backen

local-aillama-cpp-releases
6 Jun 2026
Local Ai

b9519

DGX agent

b9519 is a release build identifier for llama.cpp, an open-source C/C++ project that enables LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware . This specif

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9521

DGX agent

b9521 is an intermediate build release of llama.cpp, an open-source C/C++ library for large language model inference. The release is part of the continuous development cycle following other recent bui

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9523

DGX agent

b9523 is a release build of llama.cpp, an open-source tool for LLM inference in C/C++ that enables language model execution with minimal setup on a wide range of hardware locally and in the cloud. The

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9524

DGX agent

B9524 is a release of llama.cpp, an LLM inference project in C/C++ . This specific build number represents a snapshot from the llama.cpp GitHub releases, which are regularly published to track increme

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9528

DGX agent

Release b9528 of llama.cpp was published on June 5, 2026. This build includes a UI update to run npm install when package-lock.json is newer than node_modules, and provides pre-compiled binaries for m

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9529

DGX agent

b9529 is a release tag for llama.cpp, a C/C++ implementation of LLM inference that enables running large language models locally on consumer hardware. This specific release represents a particular ver

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9530

DGX agent

B9530 is a release version of llama.cpp, a tool for LLM inference in C/C++ . It allows users to run large language models on everyday consumer hardware without expensive GPUs or cloud infrastructure .

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9533

DGX agent

Based on the available search results, I cannot locate specific details about the b9533 release. However, b9533 is a release from the llama.cpp project, which provides LLM inference in C/C++ . The lla

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9535

DGX agent

The search did not return specific details about the b9535 release. Based on the available information, b9535 is an intermediate build version from the llama.cpp project, which is a C/C++ implementati

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9536

DGX agent

The search results don't contain specific information about release b9536, but I found related information about nearby releases in the llama.cpp project. Based on the pattern visible in the search re

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9500

DGX agent

B9500 is a release of llama.cpp that includes a Metal backend optimization reducing reset heartbeat timing from 500ms to 5ms . The release provides compiled binaries across multiple platforms includin

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9503

DGX agent

Release b9503 of llama.cpp fixes multimodal (mtmd) support by handling Gemma 4 audio projector embedding size , specifically removing the projection_dim from clip_n_mmproj_embd. This build includes pr

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9504

DGX agent

b9504 is a build of llama.cpp that includes a CMake update to skip cvector-generator and export-lora when CPU backend is disabled . The release was published on June 4, 2026, with pre-built binaries a

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9505

DGX agent

b9505 is a latest release of llama.cpp that includes a commit adding a header to tools/server/server-http.h (#24089). The release provides precompiled binaries for multiple platforms including macOS,

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9509

DGX agent

Release b9509 is a build version of llama.cpp, which provides LLM inference in C/C++ . As a specific build identifier in the llama.cpp project's release history, it likely contains bug fixes, performa

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9512

DGX agent

b9512 is a release build of llama.cpp, which provides LLM inference in C/C++ . The release uses a version numbering system with 'b' prefixes for intermediate builds, with newer releases building upon

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9515

DGX agent

llama.cpp is a C/C++ implementation designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. Build b9515 is an interme

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9518

DGX agent

llama.cpp is a C/C++ implementation that enables LLM inference with minimal setup and high performance across diverse hardware . Build b9518 is a release version from the llama.cpp project, which serv

local-aillama-cpp-releases
4 Jun 2026
Local Ai

b9487

DGX agent

The search results show general llama.cpp information and references to other recent builds (like b9484), but the specific details for b9487 were not clearly accessible. Based on the context from llam

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9488

DGX agent

Release b9488 of llama.cpp was published on June 3, 2026 , and includes support for Qwen3 SSM architectures with additions like LLM_KV_ATTENTION_RECURRENT_LAYERS . The release also fixes a bug in comm

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9489

DGX agent

Release b9489 of llama.cpp includes updates to hidden_act mapping in llama-model.cpp, additions of granite embedding multilingual R2 models, and support for setting hidden_activation in GGUF files. Th

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9491

DGX agent

b9491 is a release of llama.cpp , a C/C++ implementation of large language model inference that enables running LLMs locally with minimal dependencies. This release likely contains bug fixes, feature

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9493

DGX agent

B9493 is a release of llama.cpp, an LLM inference framework in C/C++ . This release includes updates to model support, such as centralized hidden activation mappings and additions for granite embeddin

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9494

DGX agent

The search results do not contain specific details about the b9494 release. Based on the context from the llama.cpp GitHub releases page, b9494 is an intermediate build release of llama.cpp, a C/C++ p

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9495

DGX agent

The search didn't return specific information about release b9495. Based on the available information about llama.cpp releases, here's a summary: llama.cpp enables LLM inference with minimal setup and

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9496

DGX agent

The search results do not contain specific information about release b9496. Based on the available context, b9496 is a build/release version of llama.cpp, a C/C++ implementation for efficient LLM infe

local-aillama-cpp-releases
3 Jun 2026
Local Ai

b9466

DGX agent

B9466 is a build release in the llama.cpp project that includes fixes and improvements to speculative decoding functionality, specifically addressing n_outputs_max issues and extracting helper functio

local-aillama-cpp-releases
2 Jun 2026
Local Ai

b9467

DGX agent

The search results show recent llama.cpp releases but do not contain specific information about release b9467. Based on the context of llama.cpp releases and the GitHub repository structure, b9467 is

local-aillama-cpp-releases
2 Jun 2026
← Previous
1…56789…14
Next →