AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
525 results
Model Releases

b10150

DGX agent

ggml : adjust logic for offloading ops to weight's backend (#25832) ggml : adjust logic for offloading ops to weight's backend llama : dsv4 graph fixes Website: https://llama.app macOS/iOS: macOS Appl

model-releasesllama-cpp-releases
27 Jul 2026
Model Releases

b10151

DGX agent

sycl(build): parallelize ocloc invocations (#25903) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
llama-cpp-releases
27 Jul 2026
Model Releases

b10152

DGX agent

fit : count nextn (MTP) blocks in n_gpu_layers so front layers stay on GPU (#26177) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISA

model-releasesllama-cpp-releases
27 Jul 2026
Model Releases

b10154

DGX agent

common : add common_print_available_devices() (#26170) Signed-off-by: Adrien Gallouët angt@huggingface.co Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64,

model-releasesllama-cpp-releases
27 Jul 2026
Model Releases

b10141

DGX agent

mtmd: fix android build (#26150) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubunt

model-releasesllama-cpp-releases
26 Jul 2026
Model Releases

b10103

DGX agent

metal : add f16 type support to leaky relu (#25981) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10105

DGX agent

args: refactor mlock/mmap/directio into load-mode (#20834) args: overhaul mmap/mlock/dio into single arg Signed-off-by: Aaron Teo aaron.teo1@ibm.com docs: update docs with llama-gen-docs Signed-off-by

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10106

DGX agent

CUDA: fix external compilation of q1_0 MMQ (#25778) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFra

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10107

DGX agent

hexagon: fix Windows crash when op_poll is enabled (#26029) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) i

model-releasesllama-cpp-releases
24 Jul 2026
Model Releases

b10092

DGX agent

ggml: enable PowerPC backend variants on AIX (#25983) ggml: enable PowerPC backend variants on AIX Allow the PowerPC CPU backend variants to be built on AIX by extending the platform check in the CMak

model-releasesllama-cpp-releases
23 Jul 2026
Model Releases

b10093

DGX agent

Fix DeepSeek4 crafted template (#25414) chat: fix DS4 template to explicitly follow reference behavior Support DeepSeekv4 flag (drop_reasoning). fix: hook DS3.2 parser for DS4 as well fix: add tool re

model-releasesllama-cpp-releases
23 Jul 2026
Model Releases

b10098

DGX agent

hexagon: activation ops update (#25974) hex-geglu: optimized all-in-one geglu microkernel hex-geglu: enable non-contiguous src and strided DMA hex-act: enable non-contiguous srs and strided DMA for re

model-releasesllama-cpp-releases
23 Jul 2026
Model Releases

b10099

DGX agent

CUDA: Improve NVFP4 W4A4 activation quantization (#25730) Squash history before conflict-resolution during rebase on master WIP commit Add 32-byte loads, restore per-block amax Use nvfp4x4 intrinsic w

model-releasesllama-cpp-releases
23 Jul 2026
Model Releases

b10083

DGX agent

cuda: add sqrt_softplus in topk-moe for dsv4 (#25896) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCF

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10084

DGX agent

hexagon: check tensor type when reusing descriptors (#25968) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64)

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10085

DGX agent

mtmd : use align_corners for qwen3vl vision position embedding interpolation (#25781) The Qwen3-VL learned position embedding is interpolated to the runtime patch grid with the default bilinear+antial

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10087

DGX agent

Add support for Laguna XS.2 & M.1 (#25165) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Li

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10088

DGX agent

llama-arch: fix DeepSeek4 APE tensor op (#25945) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10089

DGX agent

cuda: GET_ROWS quants (#25962) cuda: add k-quant support to GET_ROWS Device-side embedding lookups require GET_ROWS to handle the k-quants used by common GGUF recipes (Q4_K_M stores token_embd as q6_K

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10090

DGX agent

webgpu : add CONV_2D_DW (depthwise conv2d) kernel (#25847) webgpu : add CONV_2D_DW (depthwise conv2d) kernel Implement GGML_OP_CONV_2D_DW for the WebGPU backend, ported from the Vulkan backend's conv2

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10091

DGX agent

ci : fix SYCL package shared library lookup (#25987) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFr

model-releasesllama-cpp-releases
22 Jul 2026
Model Releases

b10036

DGX agent

opencl: disable FA and MoE weights repack to work around compiler issues for Adreno 850 GPU (#25745) opencl: workaround for A850 compiler compat opencl: fix DX compiler version parsing and cleanup Co-

model-releasesllama-cpp-releases
16 Jul 2026
Model Releases

b10038

DGX agent

ci : add official website link to release notes (#25728) Assisted-by: pi:llama.cpp/Qwen3.6-27B Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI en

model-releasesllama-cpp-releases
16 Jul 2026
Local Ai

b9934

DGX agent

llama.cpp is an open-source software library that performs inference on various large language models , and b9934 represents a specific build release in the project's continuous versioning system. The

local-aillama-cpp-releases
9 Jul 2026
Local Ai

b9908

DGX agent

b9908 is a build-tagged release from llama.cpp , the open-source C/C++ inference engine for large language models. llama.cpp is an open-source software library that performs inference on various large

local-aillama-cpp-releases
8 Jul 2026
Local Ai

b9932

DGX agent

B9932 is a continuous build-tagged release from the llama.cpp project , an open-source C/C++ inference engine for running large language models locally. llama.cpp performs inference on various large l

local-aillama-cpp-releases
8 Jul 2026
Local Ai

b9839

DGX agent

Build b9839 of llama.cpp adds offline mode support to the llama download command for checking cached models without network access, and fixes a use-after-free bug in the URL-task callback. llama.cpp i

local-aillama-cpp-releases
29 Jun 2026
Local Ai

b9832

DGX agent

B9832 is a build release of llama.cpp from the ggml-org project , which is a free and open-source tool that allows you to run AI models locally on Windows, Linux and macOS . The release likely contain

local-aillama-cpp-releases
28 Jun 2026
Local Ai

b9825

DGX agent

b9825 is a release of llama.cpp, the C/C++ implementation for large language model inference . This build represents an intermediate version in the ongoing development of llama.cpp, part of the ggml-o

local-aillama-cpp-releases
27 Jun 2026
Local Ai

b9810

DGX agent

b9810 is a release build of llama.cpp, an open-source software library that performs inference on various large language models such as Llama. The build identifier follows the project's versioning sch

local-aillama-cpp-releases
26 Jun 2026
Local Ai

b9816

DGX agent

B9816 is a release build number for llama.cpp, an open-source project that enables efficient large language model inference in C/C++ on consumer hardware. This release likely contains bug fixes, perfo

local-aillama-cpp-releases
26 Jun 2026
Local Ai

b9585

DGX agent

Release b9585 of llama.cpp fixes granite speech model inference by applying embedding scale when deepstack is not used . The release was published on June 9, 2026, and represents a bug fix within the

local-aillama-cpp-releases
9 Jun 2026
Local Ai

b9548

DGX agent

Release b9548 of llama.cpp includes a spec fix for vocabulary compatibility checking , addressing issues related to model format validation. This build provides compiled binaries for multiple platform

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9549

DGX agent

Release b9549 of llama.cpp adds support for the Gemma4 MTP model architecture . The release includes pre-built binaries for multiple platforms including macOS, Linux, Android, and Windows with various

local-aillama-cpp-releases
7 Jun 2026
Local Ai

b9542

DGX agent

B9542 is a build/release version of llama.cpp, the C/C++ implementation of large language model inference designed to enable LLM inference with minimal setup and state-of-the-art performance on variou

local-aillama-cpp-releases
6 Jun 2026
Local Ai

b9529

DGX agent

b9529 is a release tag for llama.cpp, a C/C++ implementation of LLM inference that enables running large language models locally on consumer hardware. This specific release represents a particular ver

local-aillama-cpp-releases
5 Jun 2026
Local Ai

b9491

DGX agent

b9491 is a release of llama.cpp , a C/C++ implementation of large language model inference that enables running LLMs locally with minimal dependencies. This release likely contains bug fixes, feature

local-aillama-cpp-releases
3 Jun 2026
Local Ai

v0.30.2

DGX agent

Ollama v0.30.2 is a patch release from the 0.30 series, which features improved compatibility and performance using llama.cpp, augmented MLX engine support on Apple Silicon, and broader model support

local-aiollama-releases
3 Jun 2026
Model Releases

v0.30.4-rc0: Kill llama-server during Windows cleanup (#16458)

DGX agent

This release candidate fixes a Windows-specific issue where the llama-server process wasn't being properly terminated during cleanup operations. The fix addresses GitHub issue #16458 and improves the

model-releasesollama-releases
3 Jun 2026
Local Ai

b9471

DGX agent

B9471 is a build release of llama.cpp, an open-source C/C++ implementation for LLM inference that enables running large language models on consumer hardware with optimized performance. This intermedia

local-aillama-cpp-releases
2 Jun 2026
Local Ai

b9431

DGX agent

b9431 is a release commit of llama.cpp, an open-source C/C++ implementation for running large language model inference efficiently on consumer hardware. Based on the source material, this release like

local-aillama-cpp-releases
30 May 2026
Local Ai

b9436

DGX agent

B9436 is an intermediate build release of llama.cpp, the open-source C/C++ implementation for running large language model inference on consumer hardware. This build typically includes updates to mode

local-aillama-cpp-releases
30 May 2026
Local Ai

b9311

DGX agent

b9311 is a release build version of llama.cpp, an open-source C/C++ framework for large language model inference. The project enables LLM inference in C/C++ , offering optimized performance across var

local-aillama-cpp-releases
25 May 2026
Local Ai

b9296

DGX agent

b9296 is a build release of llama.cpp , the C/C++ implementation for efficient large language model inference. As an intermediate build in the llama.cpp release cycle, it likely includes recent bug fi

local-aillama-cpp-releases
23 May 2026
Local Ai

b9292

DGX agent

Release b9292 of llama.cpp fixes a memory leak in the server context where speculative decoder, draft context, and draft model were not properly freed during destroy(), causing VRAM leaks on sleep/res

local-aillama-cpp-releases
22 May 2026
Local Ai

b9264

DGX agent

b9264 is a llama.cpp release that includes improvements to HunyuanVL model support, merging HunyuanOCR functionality and fixing vision precision issues. This build represents an intermediate developme

local-aillama-cpp-releases
21 May 2026
Local Ai

b9270

DGX agent

Release b9270 adds support for the HybridDNATokenizer used by the Carbon-3B model family, implementing a new BPE pre-type for tokenizing DNA sequences. The tokenizer handles DNA k-mers with fixed 6-me

local-aillama-cpp-releases
21 May 2026
Local Ai

b9134

DGX agent

B9134 is a recent build release of llama.cpp, an open-source C/C++ library for efficient large language model inference. The release was tagged on May 13, 2026, representing an ongoing development ite

local-aillama-cpp-releases
13 May 2026
← Previous
1…34567…11
Next →