AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
762 results
7 Jul 2026

b9895

Local AiDGX agent

Release b9895 of llama.cpp includes a fix for speculative inference out-of-bounds read in ngram-map on prompt shrink, with ~2x performance gains in PP_Speed for FP32, Q4_0 and Q8_0 models. The release

b9902

Local AiDGX agent

Release b9902 of llama.cpp was released on July 7, 2026 , and includes support for SYCL operations including cross_entropy_loss and cross_entropy_loss_back . The release provides builds across multipl

v0.31.2

Local AiDGX agent

Ollama v0.31.2-rc1 is a pre-release version released on July 6, 2026. This release includes CI improvements to avoid unbounded parallelism, fixes for CUDA toolkit lookup, updates to cloud documentatio

v0.31.2-rc2: llm: allow iGPU mmproj offload with fit padding (#16996)

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai
DGX agent

This release candidate introduces support for offloading the image GPU projection (mmproj) to an integrated GPU when using fit padding in Ollama's LLM processing, addressing technical improvements for

6 Jul 2026

b9879

Local AiDGX agent

b9879 is a release build identifier in the llama.cpp project, a C/C++ implementation of Meta's LLaMA language models. This release likely contains bug fixes, performance improvements, and feature upda

b9881

Local AiDGX agent

b9881 is a release of llama.cpp, an LLM inference tool written in C/C++. This release likely includes bug fixes, performance improvements, and platform support enhancements for running large language

b9884

Local AiDGX agent

B9884 is a llama.cpp release that addresses a Vulkan 32-bit integer overflow fix in CEIL_DIV . Released on July 6, 2026 , the build also includes platform-specific binaries for macOS, Linux, Android,

b9885

Local AiDGX agent

B9885 is a build-tagged release of llama.cpp, an open-source C/C++ inference engine for running large language models locally. The project does not use traditional semantic versions; instead it ships

b9886

Local AiDGX agent

Release b9886 of llama.cpp addresses a bug fix for K/V rotation input handling in attention mechanisms, specifically when buffers are unallocated during DFlash speculative decoding's KV-injection pass

b9891

Local AiDGX agent

b9891 is a build release of llama.cpp, the open-source C/C++ inference engine for running large language models locally. This release enables LLM inference with minimal setup and state-of-the-art perf

b9892

Local AiDGX agent

Release b9892 is a version identifier for llama.cpp, an open-source C/C++ framework for running large language model inference on consumer hardware. This specific build (b9892) represents a snapshot o

v0.31.2-rc0

Local AiDGX agent

v0.31.2-rc0 is a release candidate for Ollama that removes the OLLAMA_EXPERIMENT=client2 experimental flag and updates the MLX engine to the latest version . The release includes contributions from mu

5 Jul 2026

b9876

Local AiDGX agent

llama.cpp b9876 is a build-tagged release from the llama.cpp project , a pure C/C++ implementation of large language model inference . The project enables LLM inference with minimal setup and state-of

4 Jul 2026

b9871

Local AiDGX agent

The search results don't contain specific details about release b9871. Based on the pattern evident in the search results and the context, here's the summary: B9871 is a release of llama.cpp, an open-

b9873

Local AiDGX agent

Release b9873 is a version of llama.cpp, a C/C++ project designed to enable LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware locally and in the cloud. The

3 Jul 2026

b9862

Local AiDGX agent

B9862 is a continuous build-tagged release from llama.cpp , the open-source C/C++ LLM inference engine. Llama.cpp is the inference engine that powers most of the local-AI ecosystem, including tools li

b9864

Local AiDGX agent

llama.cpp is a tool for LLM inference in C/C++ and b9864 is a release version tag from the ggml-org/llama.cpp GitHub repository. Based on the release numbering pattern and project scope, this release

b9870

Local AiDGX agent

b9870 is a llama.cpp release dated July 3, 2026 , which includes fixes for StepFun parser chat handling to address long reasoning loops . The release provides pre-built binaries for multiple platforms

1 Jul 2026

b9852

Local AiDGX agent

Release b9852 is a build version of llama.cpp, a C/C++ implementation that enables large language model inference with minimal setup and state-of-the-art performance across diverse hardware platforms

b9853

Local AiDGX agent

The search results show llama.cpp releases but do not contain specific information about release b9853. Based on the context from the llama.cpp project, b9853 is likely a development build release of

b9857

Local AiDGX agent

Based on the available search results, the specific release details for b9857 are not fully accessible, but this entry refers to a build release from the llama.cpp project. llama.cpp is a port of Face

b9858

Local AiDGX agent

Release b9858 is a continuous build-tagged release of llama.cpp , the open-source C/C++ project that enables large language model inference with minimal setup and optimized performance across diverse

b9859

Local AiDGX agent

Release b9859 of llama.cpp was released on July 1, 2026 , and includes updates to allow loading precompiled binary kernels from library in the OpenCL backend . This release continues development of th

30 Jun 2026

b9843

Local AiDGX agent

b9843 is a release of llama.cpp that reverts 'sched: reintroduce less synchronizations during split compute (#20793)' . The release includes pre-built binaries for multiple platforms including macOS,

b9844

Local AiDGX agent

llama.cpp is a C/C++ implementation for LLM inference with minimal setup and state-of-the-art performance on a wide range of hardware. Release b9844 is an intermediate build of the llama.cpp project f

b9846

Local AiDGX agent

Release build b9846 is an intermediate build of llama.cpp, an open-source C/C++ project for LLM inference. The llama.cpp project is the main playground for developing new features for the ggml library

b9847

Local AiDGX agent

B9847 is a build release in the llama.cpp project, which is an open-source tool that enables LLM inference with minimal setup on a wide range of hardware . This specific build release likely contains

b9848

Local AiDGX agent

llama.cpp uses continuous build-tagged releases rather than traditional semantic versions. Release b9848 is a specific build from the ggml-org/llama.cpp project, which is an LLM inference implementati

b9849

Local AiDGX agent

b9849 is a build release from llama.cpp, which enables LLM inference in C/C++ . The release represents an intermediate build version of the llama.cpp library, a widely-used open-source project for run

b9850

Local AiDGX agent

Release b9850 is a version of llama.cpp, an open-source software library for large language model inference developed alongside the GGML tensor library. This release likely includes updates, bug fixes

b9851

Local AiDGX agent

B9851 is a release of llama.cpp, an LLM inference project written in C/C++ . The release represents a version update in the llama.cpp development timeline maintained on GitHub. This build tag typicall

29 Jun 2026

b9839

Local AiDGX agent

Build b9839 of llama.cpp adds offline mode support to the llama download command for checking cached models without network access, and fixes a use-after-free bug in the URL-task callback. llama.cpp i

b9840

Local AiDGX agent

b9840 is a release of llama.cpp, a tool for LLM inference in C/C++. The release enables running AI models locally on Windows, Linux, and macOS. This specific build tag likely contains bug fixes, perfo

v0.30.12

Local AiDGX agent

v0.30.12 is a release candidate (rc0) that fixes a gemma4:12b floating point exception crash on x86, CUDA, Linux, and Windows systems. It includes improvements to ollama launch for Hermes Desktop, all

28 Jun 2026

b9829

Local AiDGX agent

Build b9829 is an intermediate release of llama.cpp, the open-source C/C++ project that enables users to run large language models on consumer hardware without expensive GPUs or cloud infrastructure.

b9830

Local AiDGX agent

llama.cpp b9830 is a release of the open-source LLM inference project enabling local language model execution with minimal setup . The release likely includes backend improvements and optimizations fo

b9831

Local AiDGX agent

Release b9831 of llama.cpp includes backend detection improvements and synchronization enhancements, particularly for async CUDA copies and Vulkan backend operations. Llama.cpp is designed to enable L

b9832

Local AiDGX agent

B9832 is a build release of llama.cpp from the ggml-org project , which is a free and open-source tool that allows you to run AI models locally on Windows, Linux and macOS . The release likely contain

b9833

Local AiDGX agent

Release b9833 is a build of llama.cpp, an open-source project for LLM inference in C/C++. This release represents an incremental update within the llama.cpp development cycle, following the standard r

27 Jun 2026

b9822

Local AiDGX agent

b9822 is a release of llama.cpp, a project for LLM inference in C/C++ . As a build version in the ongoing development sequence of the llama.cpp project, this release likely includes bug fixes, perform

b9824

Local AiDGX agent

The search results show information about other llama.cpp releases and the project generally, but do not contain specific details about release b9824. Based on the available information about llama.cp

b9825

Local AiDGX agent

b9825 is a release of llama.cpp, the C/C++ implementation for large language model inference . This build represents an intermediate version in the ongoing development of llama.cpp, part of the ggml-o

b9826

Local AiDGX agent

B9826 is a release of llama.cpp, an LLM inference project written in C/C++ . llama.cpp enables users to run large language models on consumer hardware without expensive GPUs or cloud infrastructure .

b9827

Local AiDGX agent

Release b9827 of llama.cpp, released on June 27, 2026, added a cudaMemcpy2DAsync fast path to ggml_cuda_cpy for improved CUDA tensor copying performance. When tensors are not fully contiguous but each

b9828

Local AiDGX agent

B9828 is a llama.cpp release featuring OpenCL flash attention improvements, including reworked FA kernels for f16 and f32, prefill prepass kernels, and FA kernels for q4_0 and q8_0 quantization format

26 Jun 2026

b9810

Local AiDGX agent

b9810 is a release build of llama.cpp, an open-source software library that performs inference on various large language models such as Llama. The build identifier follows the project's versioning sch

b9811

Local AiDGX agent

b9811 is a release of llama.cpp , an open-source C/C++ project for running large language model inference. The release likely includes performance improvements, bug fixes, and optimizations to the cor

b9814

Local AiDGX agent

Based on available search results, I cannot find specific details about the b9814 release. However, b9814 is a build version identifier from the ggml-org/llama.cpp repository, which is a C/C++ impleme

b9816

Local AiDGX agent

B9816 is a release build number for llama.cpp, an open-source project that enables efficient large language model inference in C/C++ on consumer hardware. This release likely contains bug fixes, perfo

b9820

Local AiDGX agent

Release b9820 of llama.cpp introduces scheduler optimizations to reduce synchronizations during split compute and improves CUDA performance with fewer synchronizations between tokens. The update inclu

b9821

Local AiDGX agent

b9821 is the latest release of llama.cpp, a C/C++ implementation of LLM inference, which includes updates to allow the application to support --version, --licenses, and --help flags . Pre-built binari

25 Jun 2026

b9786

Local AiDGX agent

b9786 is a release of llama.cpp, a tool for LLM inference in C/C++ . Llama.cpp is a free and open-source tool that allows users to run AI models locally on Windows, Linux, and macOS . The b9786 releas

b9787

Local AiDGX agent

B9787 is a build release of llama.cpp, an open-source C/C++ implementation for LLM inference created by Georgi Gerganov. The llama.cpp project enables large language model inference with minimal setup

b9789

Local AiDGX agent

Release b9789 of llama.cpp includes a fix for quantizing mixture-of-experts models with MTP (multi-token prediction) . Binaries are provided for multiple platforms including macOS, Linux, Android, and

Evaluating performance and efficiency of the GitHub Copilot agentic harness across models and tasks

AgentsDGX agent

Explore how the GitHub Copilot agentic harness delivers strong results across multiple benchmarks and leading token efficiency, while maintaining flexibility to choose among more than 20 models. The p

v0.30.11

Local AiDGX agent

Ollama v0.30.11 is a release candidate from the v0.30 series (June 2026) , which pairs the new MLX engine on Apple Silicon with continued llama.cpp improvements for enhanced compatibility across Mac a

24 Jun 2026

b9777

Local AiDGX agent

The search results do not contain specific information about the b9777 release. Based on the available information about llama.cpp releases, b9777 is an intermediate build release from the llama.cpp p

b9784

Local AiDGX agent

Build b9784 is a release of llama.cpp, a C/C++ project that enables large language model inference with minimal setup on a wide range of hardware. The release includes pre-compiled binaries for multip

23 Jun 2026

b9769

Local AiDGX agent

The search results do not contain specific details about the b9769 release. Based on the available information about llama.cpp and its release patterns, here is a knowledge base entry: Release b9769 o

b9770

Local AiDGX agent

Release b9770 of llama.cpp addresses server functionality by fixing remote preset handling and adding tests (PR #24938). The release includes pre-built binaries for multiple platforms including macOS,

← Previous
1…34567…13
Next →