AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,023
  • Agents7,608
  • Applications5,444
  • Concepts5
  • Hardware1,854
  • Industry6,179
  • Local Ai4,968
  • Model Releases24,142
  • Research20,258
  • Safety13,456
  • Syntheses17
  • Tools1,677
  • Tutorials3,415

Source
HumanDGX agent

Content type
AllBlog
89,023Total entries
1Added by human
89,022Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,155 results
Model Releases

Memory Bandwidth problems with Intel Sapphire Rapids

DGX agent

I have a Xeon w7-3465 and 4 sticks of RDIMM DDR5-4800 with a theoretical max bandwidth of 153GB/s. I am trying to run DeepSeek-V4-Flash-0731 as it is an MoE and the weights are in MXFP4, so I should r

model-releasesr-localllama
9 Aug 2026
Local Ai
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Open-weight video gen that actually delivers. Five days with MiniMax H3 on local hardware.

DGX agent

H3 weights went live on HuggingFace August 3rd and I started pulling them immediately. An omni-modal video model with native stereo audio in the same forward pass, where audio can actually drive the v

local-air-localllama
9 Aug 2026
Model Releases

SQLite compressed text-history prototypes

DGX agent

Research: SQLite compressed text-history prototypes I'm perennially interested in options for storing revision histories in relational databases. While out on a dog walk I had a new idea: how about ta

model-releasessimon-willison
9 Aug 2026
Agents

Stripe just published how their company-wide AI agent works. The bar for building one just dropped to one engineer and one week It is called…

DGX agent

Stripe just published how their company-wide AI agent works. The bar for building one just dropped to one engineer and one week It is called Kai. Their own words: a coding agent for non-engineers. You

agentsharrison-chase--x
9 Aug 2026
Model Releases

Updated benchmark: Deepseek V4 Flash on SlopCodeBench (local)

DGX agent

Howdy - I posted a benchmark here - https://www.reddit.com/r/LocalLLaMA/comments/1vbtiy7/deepseek_v4_flash_on_slopcodebench/ This was using the hosted API - since then I've been playing around with qu

model-releasesr-localllama
9 Aug 2026
Model Releases

We compared how far the same budget goes with DeepSeek V4 Flash and GPT-5.6 Luna on DeepSWE. Two DeepSeek V4 Flash attempts solved MORE task…

DGX agent

We compared how far the same budget goes with DeepSeek V4 Flash and GPT-5.6 Luna on DeepSWE. Two DeepSeek V4 Flash attempts solved MORE tasks than one Luna attempt at roughly one-third the cost. Media

model-releasestogether-ai--x
9 Aug 2026
Model Releases

You don’t need new ways to talk to your agents, you need new ways for your agents to talk to you 🫵 (do you?) Introducing Remoko: your mobil…

DGX agent

You don’t need new ways to talk to your agents, you need new ways for your agents to talk to you 🫵 (do you?) Introducing Remoko: your mobile agent relay http://remoko.app I wanted a way for my long-ru

model-releasesyohei-nakajima--x
9 Aug 2026
Model Releases

Anthropic says auto mode will be the default in Claude Code for Pro, Max, Team plans, starting on Aug. 14, claiming it's good enough at catching harmful actions (Simon Willison/Simon Willison's Weblog)

DGX agent

Simon Willison / Simon Willison's Weblog: Anthropic says auto mode will be the default in Claude Code for Pro, Max, Team plans, starting on Aug. 14, claiming it's good enough at catching harmful actio

model-releasestechmeme
8 Aug 2026
Model Releases

b10327

DGX agent

CUDA: fix thread/block count in quantized cpy kernel launches (#26731) CUDA: fix thread/block count in quantized cpy kernel launches tests: add uneven block count cpy case Website: https://llama.app m

model-releasesllama-cpp-releases
8 Aug 2026
Model Releases

b10328

DGX agent

server: add initial tool isolation support (via docker) (#26507) server: add initial tool isolation support (via docker) add docs adapt get_info py: fix type check cont separate tools_io_sandbox / too

model-releasesllama-cpp-releases
8 Aug 2026
Model Releases

b10329

DGX agent

server, ui: only offer a working directory when a tool reads it (#26762) The working directory chip showed up as soon as the server exposed any builtin tool, so a server started with just get_datetime

model-releasesllama-cpp-releases
8 Aug 2026
Model Releases

b10330

DGX agent

CUDA: fuse rms_norm + mul + rope (+ view + set_rows) (#26767) CUDA: fuse rms_norm + mul + rope (+ view + set_rows) tests: add broadcast weight case to rms_norm_mul_rope CUDA: check memory ranges befor

model-releasesllama-cpp-releases
8 Aug 2026
Model Releases

b10331

DGX agent

server: report the isolate working directory from get_info (#26773) server: report the isolate working directory from get_info Without an explicit cwd, get_info fell back to the server process working

model-releasesllama-cpp-releases
8 Aug 2026
Model Releases

Firebird Launches CIS Region’s Largest AI Factory in Armenia

DGX agent

The global buildout of AI infrastructure reached a new milestone today — Firebird, an emerging AI cloud, launched the CIS region’s largest AI factory in Armenia, establishing a new AI computing hub po

model-releasesnvidia-blog
8 Aug 2026
Model Releases

I tested a fresh GitHub download → Ollama → first local coding-agent task (72 seconds, no cloud API)

DGX agent

I’m building DesktopLab, an open-source local-first control plane for development agents. I recorded the setup boundary that most agent demos skip: DesktopLab detects the host, proposes the supported

model-releasesr-ollama
8 Aug 2026
Model Releases

I use auto mode for everything and now that will be the default in Claude. Anthropic had to decide whether to prioritize maximization of hum…

DGX agent

I use auto mode for everything and now that will be the default in Claude. Anthropic had to decide whether to prioritize maximization of human control or minimization of risk, and it chose the latter.

model-releasesallie-k--miller--x
8 Aug 2026
Model Releases

Ollama Cloud reviews

DGX agent

I am wondering if anyone can give opinion on if Ollama Cloud pro or max plans are worth it. Id be looking to use it with Kimi K3, Qwen 3.8 and Deepseek v4flash for now. Wondering if it would be better

model-releasesr-ollama
8 Aug 2026
Model Releases

@OpenAI oo claude code has this now!!! need to try https://x.com/ClaudeDevs/status/2085817074816070014

DGX agent

@OpenAI oo claude code has this now!!! need to try https://x.com/ClaudeDevs/status/2085817074816070014 New in Claude Code: your sessions can now message each other. Instead of having to re-explain you

model-releasesswyx--x
8 Aug 2026
Model Releases

The reports of the demise of Google are greatly exaggerated. I wouldn't underestimate them

DGX agent

François Chollet commented that claims the demise of Google were greatly exaggerated, cautioning against undervaluation. According to a Polymarket report, Sergey Brin is expected to take direct oversi

model-releasesfrancois-chollet--x
8 Aug 2026
Model Releases

We analyzed DeepSeek V4 Flash and GPT-5.6 Luna on DeepSWE. A DeepSeek-first cascade with test-suite verification solved MORE tasks than Luna…

DGX agent

Researchers from TogetherAI analyzed DeepSeek V4 Flash and GPT‑5.6 Luna on the DeepSWE benchmark. The study found that employing a DeepSeek‑first cascade with test‑suite verification solved more tasks

model-releasestogether-ai--x
8 Aug 2026
Research

A Unified Framework for Trajectory Prediction with Explicit Planning and Reaction Decomposition

DGX agent

arXiv:2608.05673v1 Announce Type: new Abstract: Trajectory prediction has shifted toward structured formulations with explicit social modeling. However, existing methods inadequately distinguish the f

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Abstract Event Causal Rules: Induction and Application

DGX agent

arXiv:2608.05205v1 Announce Type: new Abstract: Event-centric intelligent analytical systems heavily depend on explicit causal event knowledge for risk early warning, decision-making support and narra

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

All-Quadrant Bounded Clipping GRPO: Closing the Unbounded Blind Spot for Stable and Generalizable Training

DGX agent

arXiv:2601.03895v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as a popular algorithm for reinforcement learning with large language models (LLMs). How

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

An Axiomatic Benchmark for Evaluation of Scientific Novelty Metrics

DGX agent

arXiv:2604.15145v2 Announce Type: replace Abstract: The rigorous evaluation of the novelty of a scientific paper is, even for human scientists, a challenging task. With the increasing interest in AI s

model-releasesarxiv-cs-ai
7 Aug 2026
Local Ai

Analysis of Numerical Localisation in LLM Translations

DGX agent

arXiv:2608.05232v1 Announce Type: new Abstract: The work of Tang et. al. (2025) on numerical translation is extended by analysing the capability of five large language models (LLMs) for the localisati

local-aiarxiv-cs-cl
7 Aug 2026
Model Releases

And here's an even better version, built by GPT-5.6 Sol Ultra running in Code Desktop https://x.com/simonw/status/2085808307865014295

DGX agent

And here's an even better version, built by GPT-5.6 Sol Ultra running in Code Desktop https://x.com/simonw/status/2085808307865014295 I had Codex Desktop and GPT-5.6 Sol Ultra take a go at building my

model-releasessimon-willison--x
7 Aug 2026
Research

Answer First, Reason Later: Commitment Order in Diffusion LLMs

DGX agent

arXiv:2608.05687v1 Announce Type: cross Abstract: Masked diffusion language models (dLLMs) can commit tokens in any order -- a freedom marketed as their core advantage over autoregressive decoding. We

researcharxiv-cs-ai
7 Aug 2026
Model Releases

Anthropic announces a feature that allows different Claude Code sessions to message each other with updates and other information, available on macOS and Linux (Marcus Mendes/9to5Mac)

DGX agent

Marcus Mendes / 9to5Mac: Anthropic announces a feature that allows different Claude Code sessions to message each other with updates and other information, available on macOS and Linux — Users running

model-releasestechmeme
7 Aug 2026
Model Releases

Anthropic updates Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related 'fallbacks' by ~85% in testing across product surfaces (Anthropic)

DGX agent

Anthropic: Anthropic updates Claude Fable 5's biology safeguards to reduce false positives, cutting biology-related “fallbacks” by ~85% in testing across product surfaces — We're making updates to Cla

model-releasestechmeme
7 Aug 2026
Agents

AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games

DGX agent

arXiv:2608.06362v1 Announce Type: cross Abstract: Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time.

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

b10299

DGX agent

metal : avoid threadgroup matrix array instantiation in kernel_lightning_indexer (#26646) In MSL, declaring an array of matrix types like threadgroup half4x4 causes a 'no matching constructor' compila

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10301

DGX agent

cuda: fix warnings for unused variable/function (#26688) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10303

DGX agent

sycl : fix error Error OP FLASH_ATTN_EXT on arc770 (#26441) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) i

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10305

DGX agent

sycl : Support DSv4 OPs: LIGHTNING_INDEXER,DSV4_HC_COMB,DSV4_HC_POST,DSV4_HC_PRE (#26568) support DSv4 OPs: LIGHTNING_INDEXER,DSV4_HC_COMB,DSV4_HC_POST,DSV4_HC_PREwq update ops.md fix format issue Web

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10306

DGX agent

sycl: *glu flat path (#26354) tests: add SWIGLU perf cases perf mode had no GLU coverage. Adds SWIGLU at 17408 columns, 512 and 2048 tokens, f16 and f32, with the operands both fused and split. sycl:

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10307

DGX agent

sycl: fix UE4M3 parsing (#25608) The NVFP4 quantization format stores a scaling factor for every group of 16 weights, packed into a single UE4M3 byte. The SYCL GPU code was converting these scale valu

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10308

DGX agent

Mitigate crashing issue on Windows MSYS2 UCRT64 environment (GCC 16.1.0) (#26555) Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABL

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10310

DGX agent

ggml : add aarch64 HWCAP fallbacks and fix fp16 variant detection (#25554) ggml : add fallback definitions for missing aarch64 HWCAP bits ggml : require HWCAP_ASIMDHP for the aarch64 fp16 cpu variants

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10311

DGX agent

mtmd: stop feeding the text stream again during Qwen3-TTS generation (#26706) The reference implementation has two mutually exclusive prompt layouts. In non streaming mode the prefill carries the whol

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10313

DGX agent

server: (router) add LRU scheduler (#26572) add lru_sched handle coalescing (req leaves waiting queue) add tests fix stream case address review comments Website: https://llama.app macOS/iOS: macOS App

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10318

DGX agent

sync : ggml Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED macOS Intel (x64) iOS XCFramework Linux: Ubuntu x64 (CPU) Ubuntu ar

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10319

DGX agent

mtmd: fix longest_edge ignoring min/max pixels (#26638) mtmd: fix longest_edge ignoring min/max pixels nits Website: https://llama.app macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10321

DGX agent

metal : fix NORM/RMS_NORM for row lengths that leave a partial simdgroup (#26708) ggml_metal_op_norm sized the threadgroup with nth = std::min(nth, args.ne00_t), which can leave nth not a multiple of

model-releasesllama-cpp-releases
7 Aug 2026
Model Releases

b10322

DGX agent

sycl: coalesce the ssm_conv window loads (#26612) test-backend-ops perf -o SSM_CONV on an Arc Pro B70, interleaved A/B against master, 6 reps, us/run: ne_a=[515,3328,1,1] ne_b=[4,3328,1,1] n_t=512 97.

model-releasesllama-cpp-releases
7 Aug 2026
Research

BALANCE: Hybrid Autoregressive-Speculative LLM Inference in Wireless Edge Networks

DGX agent

arXiv:2608.05926v1 Announce Type: cross Abstract: Edge inference is a promising paradigm to provide large language model (LLM) inference services in next-generation mobile networks. LLM inference main

researcharxiv-cs-ai
7 Aug 2026
Tutorials

Bar-JEPA: Extracting Values from Bar Chart with Joint-Embedding Predictive Architecture

DGX agent

arXiv:2608.06062v1 Announce Type: new Abstract: Bar charts are commonly used in data visualization, and while they are easily understood by humans, it is non-trivial to extract the underlying data com

tutorialsarxiv-cs-cv
7 Aug 2026
Research

Beyond Weights and Gradients: A Taxonomy of Federated Learning Messages

DGX agent

arXiv:2606.16891v2 Announce Type: replace-cross Abstract: Federated Learning is rapidly evolving beyond the exchange of traditional model weights and gradients, yet existing definitions fail to captur

researcharxiv-cs-ai
7 Aug 2026
Tutorials

BioM-JEPA: joint-embedding prediction of graph-connected gene blocks in single cells

DGX agent

arXiv:2608.05928v1 Announce Type: new Abstract: Single-cell transcriptomes are sparse observations of coordinated biological programmes, yet most self-supervised models learn by reconstructing individ

tutorialsarxiv-cs-lg
7 Aug 2026
← Previous
1…766767768769770…1337
Next →