AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
762 results
23 Jul 2026

b10092

Model ReleasesDGX agent

ggml: enable PowerPC backend variants on AIX (#25983) ggml: enable PowerPC backend variants on AIX Allow the PowerPC CPU backend variants to be built on AIX by extending the platform check in the CMak

b10093

Model ReleasesDGX agent

Fix DeepSeek4 crafted template (#25414) chat: fix DS4 template to explicitly follow reference behavior Support DeepSeekv4 flag (drop_reasoning). fix: hook DS3.2 parser for DS4 as well fix: add tool re

b10094

Model ReleasesDGX agent

common: infer the speculative type from the draft repo sidecars (#25989) With -hfd pointing to a repo that ships mtp-/dflash-/eagle3- sidecars and no --spec-type given, the draft resolved to a full mo

b10098

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

hexagon: activation ops update (#25974) hex-geglu: optimized all-in-one geglu microkernel hex-geglu: enable non-contiguous src and strided DMA hex-act: enable non-contiguous srs and strided DMA for re

b10099

Model ReleasesDGX agent

CUDA: Improve NVFP4 W4A4 activation quantization (#25730) Squash history before conflict-resolution during rebase on master WIP commit Add 32-byte loads, restore per-block amax Use nvfp4x4 intrinsic w

v0.32.3

Model ReleasesDGX agent

What's Changed mlx update by @dhiltgen in #17332 model/parsers: finalize incomplete GLM tool calls by @dhiltgen in #17250 docs: update retirements by @mxyng in #17289 model: align Laguna with upstream

22 Jul 2026

b10083

Model ReleasesDGX agent

cuda: add sqrt_softplus in topk-moe for dsv4 (#25896) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCF

b10084

Model ReleasesDGX agent

hexagon: check tensor type when reusing descriptors (#25968) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64)

b10085

Model ReleasesDGX agent

mtmd : use align_corners for qwen3vl vision position embedding interpolation (#25781) The Qwen3-VL learned position embedding is interpolated to the runtime patch grid with the default bilinear+antial

b10087

Model ReleasesDGX agent

Add support for Laguna XS.2 & M.1 (#25165) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Li

b10088

Model ReleasesDGX agent

llama-arch: fix DeepSeek4 APE tensor op (#25945) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramew

b10089

Model ReleasesDGX agent

cuda: GET_ROWS quants (#25962) cuda: add k-quant support to GET_ROWS Device-side embedding lookups require GET_ROWS to handle the k-quants used by common GGUF recipes (Q4_K_M stores token_embd as q6_K

b10090

Model ReleasesDGX agent

webgpu : add CONV_2D_DW (depthwise conv2d) kernel (#25847) webgpu : add CONV_2D_DW (depthwise conv2d) kernel Implement GGML_OP_CONV_2D_DW for the WebGPU backend, ported from the Vulkan backend's conv2

b10091

Model ReleasesDGX agent

ci : fix SYCL package shared library lookup (#25987) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFr

Copilot vs. raw API access: What are you actually paying for?

SafetyDGX agent

Copilot now bills usage at listed API rates. Compare direct model access with the coding workflow, policy, and harness work around it. The post Copilot vs. raw API access: What are you actually paying

v0.32.2

Model ReleasesDGX agent

What's Changed launch: keep Claude Code channels available by @hoyyeva in #17210 cmd: remove dead agent prompt wrappers by @ParthSareen in #17227 agent: reorder working directory instruction by @Parth

21 Jul 2026

How to build interactive experiences with canvases

TutorialsDGX agent

Canvases turn AI into interactive workspaces where you can visualize information, explore workflows, and take action across complex tasks. The post How to build interactive experiences with canvases a

v0.32.2-rc1: server: detect download stalls before the first byte (#17259)

Local AiDGX agent

The **v0.32.2‑rc1** release (commit 4d1b53e, signed with a GPG key) introduces server-side logic to detect download stalls that occur before any data is received. It also removes the stall‑timeout set

v0.32.2-rc3: test: revamp integration test entrpoints (#16560)

Local AiDGX agent

This refactors the existing integration tests into 3 priumary groups: fast, release, and library. It also refines some of the release tests to drop some of the older models and pick up newer models, w

20 Jul 2026

v0.32.2-rc0: cuda: add CC 10.0 for linux in CUDA v12 (#17025)

Local AiDGX agent

**Release v0.32.2‑rc0 (commit de1ce45)** adds support for Compute Capability 10.0 in the Linux “CUDA v12” preset, enabling B200‑class GPUs to use the `cuda_v12` backend even when drivers do not meet t

16 Jul 2026

b10036

Model ReleasesDGX agent

opencl: disable FA and MoE weights repack to work around compiler issues for Adreno 850 GPU (#25745) opencl: workaround for A850 compiler compat opencl: fix DX compiler version parsing and cleanup Co-

b10038

Model ReleasesDGX agent

ci : add official website link to release notes (#25728) Assisted-by: pi:llama.cpp/Qwen3.6-27B Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI en

v0.32.1

Model ReleasesDGX agent

What's Changed Improved Gemma 4 tool calling and multi-turn reasoning, including more reliable tool-response continuations Fixed a recurrent MLX model cache leak that could increase memory use across

15 Jul 2026

b10016

Model ReleasesDGX agent

[SYCL] Flash Attention with XMX engine via oneDNN (#25222) [SYCL] F16 (default) Flash Attention with XMX engine via oneDNN graph API; Qwen3.6-27b-Q8_0 prefill speed up x1.21 at p=512 and x4.26 at p=80

b10017

Model ReleasesDGX agent

sycl: Increase minimum buffer size for USM system allocations (#25525) Raise the threshold for minimum buffer size from 1 GiB to 4 GiB, based on real-world experiments of overcommitting device memory

b10021

Model ReleasesDGX agent

DeepseekV4: reduce graph splits (#25702) macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu

14 Jul 2026

b10011

Model ReleasesDGX agent

server : refactor prompt cache state ownership (#25649) server : clear checkpoints upon prompt clear server : move the prompt state data to the server_prompt_cache Assisted-by: pi:llama.cpp/Qwen3.6-27

v0.32.0

Model ReleasesDGX agent

What's Changed New interactive agent experience: running ollama now launches an agent to help you code and delegate work ❯ ollama Ollama 0.32.0 ▸ Chat, Code, & Work (glm-5.2:cloud) Chat with models, c

10 Jul 2026

b9947

Local AiDGX agent

The search did not return the specific details of the b9947 release. Based on the repository and context available: B9947 is a release of llama.cpp, which enables LLM inference in C/C++ . The release

b9948

Local AiDGX agent

b9948 is a build release of llama.cpp, an open-source C/C++ project for efficient large language model inference. The project ships continuous build-tagged releases using a build numbering system rath

b9949

Local AiDGX agent

b9949 is a release build of llama.cpp, the open-source C/C++ inference engine for running large language models locally. llama.cpp releases are published frequently with build versions in the b-series

b9950

Local AiDGX agent

b9950 is a release of llama.cpp, a C/C++ tool for LLM inference . This build number represents one of the frequent incremental releases from the llama.cpp project, which follows a rapid development cy

b9956

Local AiDGX agent

Release b9956 is a build version of llama.cpp, an open-source project for LLM inference in C/C++ . The project releases frequently, with multiple releases published in a single day , and b9956 represe

b9967: server: accept null sampling params (#25538)

Local AiDGX agent

Release b9967 of llama.cpp, an LLM inference project in C/C++, includes an update to the server component allowing it to accept null sampling parameters (PR #25538). This change enables more flexible

Better tools made Copilot code review worse. Here’s how we actually improved it.

AgentsDGX agent

How migrating Copilot code review to shared Unix-style code exploration tools reduced review cost by reshaping agent workflows around pull request evidence. The post Better tools made Copilot code rev

v0.32.0-rc0

Local AiDGX agent

The search results do not contain specific information about v0.32.0-rc0. Based on the version numbering and context from Ollama's release patterns, v0.32.0-rc0 is a release candidate for Ollama, a la

9 Jul 2026

b9934

Local AiDGX agent

llama.cpp is an open-source software library that performs inference on various large language models , and b9934 represents a specific build release in the project's continuous versioning system. The

b9935

Local AiDGX agent

b9935 is a build release from the llama.cpp project, which implements LLM inference in C/C++. The project doesn't follow traditional release practices, as multiple releases can be published in a singl

b9936

Local AiDGX agent

llama.cpp b9936 is a continuous build-tagged release from the open-source llama.cpp project , which is an open-source C/C++ inference engine that powers most of the local-AI ecosystem . This release c

b9938

Local AiDGX agent

The search results do not contain specific details about release b9938. However, based on the context available, b9938 is one of the regular build releases from the llama.cpp project. Llama.cpp releas

b9940

Local AiDGX agent

Release b9940 of llama.cpp was published on July 9, 2026 , and includes changes related to llama-bench initialization parameters . The release provides pre-built binaries across multiple platforms inc

b9945

Local AiDGX agent

b9945 is a build-tagged release of llama.cpp, an open-source library that performs inference on large language models and is co-developed alongside the GGML tensor library. The llama.cpp project does

b9946

Local AiDGX agent

The search results show that b9946 is a specific release build tag from the llama.cpp project, though the exact details of that particular build are not visible in the page content retrieved. Based on

8 Jul 2026

Automating cross-repo documentation with GitHub Agentic Workflows

AgentsDGX agent

Explore how the Aspire team turns merged product changes into SME-reviewed docs pull requests, closing the gap between release and documentation. The post Automating cross-repo documentation with GitH

b9905

Local AiDGX agent

Based on the available information, b9905 is a build-tagged release from the llama.cpp project, which is an open-source C/C++ inference engine that powers most of the local-AI ecosystem . The project

b9908

Local AiDGX agent

b9908 is a build-tagged release from llama.cpp , the open-source C/C++ inference engine for large language models. llama.cpp is an open-source software library that performs inference on various large

b9909

Local AiDGX agent

llama.cpp is an open-source software library that performs inference on various large language models , and the project ships continuous build-tagged releases rather than traditional semantic versioni

b9910

Local AiDGX agent

Based on available information, b9910 is a release tag from the llama.cpp project, an open-source C/C++ implementation for running large language model inference locally on consumer hardware. Llama.cp

b9913

Local AiDGX agent

b9913 is a build-tagged release from llama.cpp, an open-source software library that performs inference on various large language models. The project does not use traditional semantic versions; instea

b9914

Local AiDGX agent

The search results don't contain specific details about the b9914 release. Based on the context of llama.cpp releases, b9914 is a build/commit version in the llama.cpp project, an open-source tool for

b9916

Local AiDGX agent

b9916 is a release of llama.cpp, an open-source C/C++ implementation for LLM inference . The release represents part of the project's rapid development cycle, with binaries available for multiple plat

b9923

Local AiDGX agent

b9923 is a release build of llama.cpp, an open-source C/C++ project for LLM inference with minimal setup and state-of-the-art performance on various hardware . The specific b9923 build includes binary

b9925

Local AiDGX agent

Release b9925 of llama.cpp is a version update for the LLM inference in C/C++ project. This build is part of the project's frequent release cycle, which can publish multiple releases in a single day ,

b9929

Local AiDGX agent

The search results don't provide specific details about release b9929. Based on the available information and the context that llama.cpp releases frequently with tagged versions, b9929 is a specific b

b9931

Local AiDGX agent

Based on available information, b9931 is a release of llama.cpp, which is an open-source project for large language model inference in C/C++. As a commit-based release from the ggml-org/llama.cpp repo

b9932

Local AiDGX agent

B9932 is a continuous build-tagged release from the llama.cpp project , an open-source C/C++ inference engine for running large language models locally. llama.cpp performs inference on various large l

b9933

Local AiDGX agent

b9933 is a continuous build-tagged release of llama.cpp , the C/C++ inference engine for running large language models locally. This release represents an incremental update in llama.cpp's development

How GitHub Copilot enables zero DNS configuration for GitHub Pages

ToolsDGX agent

Go from an empty repository to a live custom domain with HTTPS in about 14 minutes, without manually editing a single DNS record. The post How GitHub Copilot enables zero DNS configuration for GitHub

7 Jul 2026

b9893

Local AiDGX agent

b9893 is a release build of llama.cpp with Windows OpenVINO 2026.2.1 support . The llama.cpp project uses continuous build-tagged releases rather than traditional semantic versioning , making b9893 on

b9894

Local AiDGX agent

Release b9894 of llama.cpp was published on July 7, 2026 , and includes a Vulkan backend fix to check src0 type in GGML_OP_SET_ROWS to avoid failures due to unimplemented f16 support . Llama.cpp is th

← Previous
123456…13
Next →