AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,675 results
Model Releases

JuryProbe: An Empirical Consensus-Risk Diagnostic for Routing Reference-Free Factuality Judge Panels to Grounded Verification

DGX agent

arXiv:2608.20607v1 Announce Type: cross Abstract: Panels of inexpensive LLM judges increasingly make accept-or-escalate decisions. In factuality settings, accepting a claim because several reference-f

model-releasesarxiv-cs-ai
24 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Lightweight Adaptive ReduNet via Hyperspherical Manifold Learning

DGX agent

arXiv:2608.20668v1 Announce Type: cross Abstract: In recent years, a white-box neural network called ReduNet has been proposed, which employs the maximal coding rate reduction (MCR^2) principle to tra

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

llm-anthropic 0.27

DGX agent

Release: llm-anthropic 0.27 This release of the Anthropic plugin for LLM mainly provides compatibility with the recently released anthropic v1.0.0 Python library, which switches from httpx to httpx2.

model-releasessimon-willison
24 Aug 2026
Research

LTR-ICD: A Ranking-Aware Framework for Automatic ICD Coding

DGX agent

arXiv:2510.13922v2 Announce Type: replace-cross Abstract: Clinical notes contain unstructured text provided by clinicians during patient encounters. These notes are usually accompanied by a sequence o

researcharxiv-cs-cl
24 Aug 2026
Model Releases

Maximum Entropy Encoding of Energy-Weighted Spherical Moments

DGX agent

arXiv:2608.20429v1 Announce Type: cross Abstract: We study how angular energy signals composed of non-negative Monte Carlo path samples can be compressed and reconstructed for irradiance using finite

model-releasesarxiv-cs-cv
24 Aug 2026
Research

Memory Augmentation Unlocks Efficient Chain-of-Thought Reasoning

DGX agent

arXiv:2608.21265v1 Announce Type: new Abstract: Large language models often rely on Chain-of-Thought (CoT) reasoning to solve complex tasks, but verbose reasoning traces introduce substantial inferenc

researcharxiv-cs-cl
24 Aug 2026
Local Ai

NeSAM: Neuro-Symbolic Kinodynamics with Soil Adaptation for Off-Road Mobility

DGX agent

arXiv:2608.21330v1 Announce Type: new Abstract: Accurate prediction of off-road vehicle motion over deformable terrain remains challenging because sinkage, slip, and traction vary with local soil cond

local-aiarxiv-cs-ro
24 Aug 2026
Research

NeuroStrata: An Electroencephalographic Connectivity-Aware Deep Representation Learning Framework for Dynamic Brain Network Analysis of Mental Stress

DGX agent

arXiv:2608.20354v1 Announce Type: cross Abstract: This study introduces NeuroStrata, a connectivity-aware deep representation learning framework for EEG-based mental stress analysis using Time-Varying

researcharxiv-cs-ai
24 Aug 2026
Hardware

NVIDIA Vera Rubin and Blackwell Set a New Standard for Agentic AI Performance per Watt

DGX agent

NVIDIA’s new GPUs in the Vera Rubin and Blackwell families set a new benchmark for agentic‑AI performance per watt on AgentX, an open‑source inference test that models realistic multi‑step, tool‑using

hardwarenvidia-developer
24 Aug 2026
Tutorials

OccluRank: Controllable Occlusion-Aware Layout-to-Image Generation by Adding Just an Ordinal Rank

DGX agent

arXiv:2608.20932v1 Announce Type: new Abstract: Layout-to-image generation enables explicit spatial control through bounding-box layouts, yet bounding boxes specify only instance locations and cannot

tutorialsarxiv-cs-cv
24 Aug 2026
Model Releases

Planning to spend ~$100 benchmarking differnet Qwen3.8-27B quants and kv cache and looking for input before I start

DGX agent

TL;DR: I'm planning to spend around 100 on cloud GPUs to benchmark Qwen3.8-27B with a focus on questions that actually matter when running it locally: different quant levels/providers, 8-bit vs 16-bit

model-releasesr-localllama
24 Aug 2026
Model Releases

Primal Acceleration of Newton's Method

DGX agent

arXiv:2608.21359v1 Announce Type: cross Abstract: We develop a new direct accelerated Newton method for minimizing convex functions with Lipschitz continuous Hessian. The algorithm uses only primal va

model-releasesarxiv-cs-ai
24 Aug 2026
Applications

Profiling What Matters: Context-Aware Item Profiles from Large-Scale Metadata for LLM Recommenders

DGX agent

arXiv:2608.20801v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have significantly advanced reranking in recommendation, effectively leveraging item-side information remains chall

applicationsarxiv-cs-ai
24 Aug 2026
Model Releases

Provable Edge-of-Stability for Adam on a One-Dimensional Quadratic

DGX agent

arXiv:2608.20638v1 Announce Type: cross Abstract: The edge-of-stability (EoS) phenomenon of Adam has been widely observed, while its underlying dynamical mechanism is not yet fully understood. We stud

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

Qwen 27B 3.8 quants: How low can you go?

DGX agent

For the GPU poor among us: I'm curious what results you're getting with low quants of Qwen 27B 3.8. My main inference hardware is limited (Mac mini M4 24 GB), but I'm getting great results with Unslot

model-releasesr-localllama
24 Aug 2026
Model Releases

qwen38-27b-rtx3090 (https://github.com/syv-ai/qwen38-27b-rtx3090) is extremely good with deepseek harness.

DGX agent

With vision enabled I am able to run at 150k context on a single RTX 3090 and the results are just amazing. I was even able to write a gmail plugin for DeepSeek harness with locally hosted Qwen 3.8 27

model-releasesr-localllama
24 Aug 2026
Model Releases

Representation Affects Retrieval: A Case Study of Skill Discovery and Routing in a Multimodal Agent Harness

DGX agent

arXiv:2608.20389v1 Announce Type: new Abstract: A production agent harness must discover and rank, from a growing library of skills, the one most appropriate for a user's task. At small scale this sel

model-releasesarxiv-cs-ai
24 Aug 2026
Safety

Scaling Unsupervised Word Alignment to Documents via Structural Constraints

DGX agent

arXiv:2608.21023v1 Announce Type: new Abstract: Word alignment has traditionally been studied between sentences, but many cross-lingual tasks increasingly require correspondences across full documents

safetyarxiv-cs-cl
24 Aug 2026
Agents

SDAD: Spec-Driven Agentic Development for the AI-Native SDLC

DGX agent

arXiv:2608.20341v1 Announce Type: new Abstract: Frontier coding agents backed by large language models with context windows from hundreds of thousands to millions of tokens are restructuring the Softw

agentsarxiv-cs-ai
24 Aug 2026
Applications

SENTRY: Deterministic, Intelligent Risk Assessment for IT Change Management

DGX agent

arXiv:2608.21203v1 Announce Type: new Abstract: Technology change management in large financial institutions depends on risk assessments that are accurate, consistent, and auditable. In practice, many

applicationsarxiv-cs-ai
24 Aug 2026
Model Releases

Specification Portability Across LLM Development Agents: Cross-Agent Compatibility in Specification-Driven Software Migration

DGX agent

arXiv:2608.21208v1 Announce Type: cross Abstract: This paper investigates cross-agent specification portability using Oracle-to-PostgreSQL migration as a controlled software transformation task. The s

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

STCO: Conditional Neural Operators for Time-Dependent PDEs

DGX agent

arXiv:2608.20477v1 Announce Type: new Abstract: Neural operators have emerged as efficient surrogates for time-dependent physical systems governed by partial differential equations (PDEs), but their f

model-releasesarxiv-cs-ai
24 Aug 2026
Agents

Structure for Reading, Prose for Writing: Asymmetric Structural Conditioning in Multi-Agent Document Authoring

DGX agent

arXiv:2608.20786v1 Announce Type: new Abstract: Multi-agent pipelines that author formal documents must both read a requester's forms and write against them. We report a deployed tender-response syste

agentsarxiv-cs-ai
24 Aug 2026
Research

SuppreSensing: Expert-Guided Feature Recalibration and Discrepancy Augmentation for Multimodal Object Detection

DGX agent

arXiv:2608.20944v1 Announce Type: new Abstract: Multimodal object detection in remote sensing faces challenges due to semantic heterogeneity and modality-specific noise interference. To this end, we p

researcharxiv-cs-cv
24 Aug 2026
Model Releases

TaPeR: Probabilistic Recovery of Sparse Task Precedence Graphs from a Handful of Demonstrations

DGX agent

arXiv:2608.21035v1 Announce Type: new Abstract: Long-horizon manipulation tasks are often only partially ordered. For example, when assembling an electronic device, the battery and circuit board may b

model-releasesarxiv-cs-ro
24 Aug 2026
Research

Temporal Validity on Real Software Histories: Eliminating Stale-Fact Errors in Code-Assistant Memory over GitHub Fixes

DGX agent

arXiv:2608.20685v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has no model of time: when a fact changes across a coding session - a function is renamed, an endpoint moves, a d

researcharxiv-cs-ai
24 Aug 2026
Safety

This is one of the most useful writeups I have seen on keeping an LLM judge effective in production. (bookmark it) Netflix runs judges over …

DGX agent

This is one of the most useful writeups I have seen on keeping an LLM judge effective in production. (bookmark it) Netflix runs judges over hundreds of thousands of show-level recommendation explanati

safetydair-ai--x
24 Aug 2026
Model Releases

TRACE: Training-time Report-guided and Clinically Ordered Concept Editing

DGX agent

arXiv:2608.20809v1 Announce Type: cross Abstract: Breast ultrasound diagnosis relies on clinically meaningful semantic concepts, yet most deep learning methods adopt end-to-end image-to-label paradigm

model-releasesarxiv-cs-ai
24 Aug 2026
Model Releases

Tree-of-Concerns: Hierarchical Multi-Agent Debate for Unstated-Limitation Extraction in Scientific Critique

DGX agent

arXiv:2608.20777v1 Announce Type: new Abstract: As scientific literature grows and papers increasingly under-report limitations, multi-agent LLMs offer a promising approach to systematically uncover t

model-releasesarxiv-cs-cl
24 Aug 2026
Model Releases

Triangulation-Free Bundle Adjustment with Graduated Non-Convexity for Camera Pose Refinement from Coarse Priors

DGX agent

arXiv:2608.21008v1 Announce Type: new Abstract: Mobile AR frameworks attach a metric pose prior to every casual phone capture, and turning it into reconstruction-grade poses cheaply on CPU is the step

model-releasesarxiv-cs-cv
24 Aug 2026
Model Releases

When Failures Propagate: Causal Failure Attribution in Agentic Retrieval-Augmented Generation

DGX agent

arXiv:2608.20627v1 Announce Type: cross Abstract: Agentic retrieval-augmented generation (RAG) interleaves retrieval, reasoning, and answer generation across multiple hops. A retrieval error at hop 1

model-releasesarxiv-cs-ai
24 Aug 2026
Research

When Graph-JEPA Learns the Wrong Thing: Diagnosing and Repairing Category-Conditional Collapse

DGX agent

arXiv:2608.20516v1 Announce Type: new Abstract: Joint-embedding predictive architectures are selected almost universally by linear probing and effective rank. We report a case where both read healthil

researcharxiv-cs-lg
24 Aug 2026
Model Releases

When Retrieval Fails Before It Begins: Structurally Indirect Prerequisite Eviction as a Retention Failure in Agentic Memory

DGX agent

arXiv:2608.20400v1 Announce Type: new Abstract: Agentic memory under a fixed budget involves two stages: retention and retrieval. Existing retrieval-centered paradigms implicitly assume necessary evid

model-releasesarxiv-cs-ai
24 Aug 2026
Local Ai

Zero-Shot Color Image Manipulation Localization via Noise Residual Artifact Pattern Analysis

DGX agent

arXiv:2608.20558v1 Announce Type: new Abstract: Digital cameras embed device-specific artifacts into every acquired image through demosaicing, in-camera post-processing, and lossy compression. These t

local-aiarxiv-cs-cv
24 Aug 2026
Model Releases

Agent Quest now tells you when Claude Code or Codex needs you visually and with sound

DGX agent

A few weeks ago I shared Agent Quest, my open-source experiment that turns Claude Code and Codex sessions into heroes living inside a small 2D world. The original idea was mainly about making it easie

model-releasesr-localllama
23 Aug 2026
Model Releases

b10589

DGX agent

cuda : add POOL_1D support (#27573) cuda : add POOL_1D support fix: add missing trailing newline for editorconfig compliance Website: https://llama.app Attestations: https://github.com/ggml-org/llama.

model-releasesllama-cpp-releases
23 Aug 2026
Model Releases

b10590

DGX agent

vendor : update subprocess.h (#27409) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/42402532 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (a

model-releasesllama-cpp-releases
23 Aug 2026
Model Releases

b10594

DGX agent

common : skip device_info loop if it's not going to be printed (#26692) The device_info loop iterates over the discovered devices and gets the available and total memory counts. With the CUDA backend

model-releasesllama-cpp-releases
23 Aug 2026
Model Releases

b10595

DGX agent

server : add LLAMA_SERVER_SLOTS_N_DIFF (#27600) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/42423433 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple

model-releasesllama-cpp-releases
23 Aug 2026
Model Releases

b10599

DGX agent

test: move tools/parser to tests (#27548) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/42442638 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silico

model-releasesllama-cpp-releases
23 Aug 2026
Model Releases

I guess children needlessly dying of a disease that has long been fully preventable is a small price to pay for Byron Donalds to become gove…

DGX agent

I guess children needlessly dying of a disease that has long been fully preventable is a small price to pay for Byron Donalds to become governor of Florida. Brennan: 'Nearly all of Florida is below th

model-releasesanthropic--x
23 Aug 2026
Model Releases

I hosted Kimi K3 (2.8T parameters) using 8 B300s. 92 tok/s, $190 per million tokens

DGX agent

What I ran: 8x B300 on Modal, 56.79 per hour, vLLM, tensor parallel 8, native MXFP4 Cold boot ~27 min (1.56 TB load, JIT, 51 CUDA graph captures) TTFT 0.92 to 1.02 s, decode 92 tok/s steady, 83 tok/s

model-releasesr-localllama
23 Aug 2026
Model Releases

If you maintain an AGENTS.md or a CLAUDE.md, this one is worth your time. (bookmark it) Researchers traced 94K development events across 557…

DGX agent

If you maintain an AGENTS.md or a CLAUDE.md, this one is worth your time. (bookmark it) Researchers traced 94K development events across 557 agentic coding sessions, plus 690K file-level change record

model-releasesdair-ai--x
23 Aug 2026
Model Releases

Looking at a PowerColor R9700 for Qwen3.8-27B, Q4_K_XL, llama.cpp/Vulkan.

DGX agent

Hi all Looking at a PowerColor R9700 for Qwen3.8-27B, Q4, llama.cpp/Vulkan. AMD's own blog quotes 51.8 tok/s but doesn't say what context length that's at, or whether MTP=2 was holding up. Separately

model-releasesr-localllama
23 Aug 2026
Model Releases

Mozilla Killed Orbit. I Rebuilt It Locally and Privately.

DGX agent

Hey everyone! Last year, Mozilla released Orbit, an AI-powered browser summarizer hosted on a GCP server. After people started digging into the extension, they discovered things like backend endpoints

model-releasesr-ollama
23 Aug 2026
Research

Scientific terms should have precision. If we use the terms VLM, VLA, WAM in an indiscriminate fashion, as is becoming common in robotics, w…

DGX agent

Scientific terms should have precision. If we use the terms VLM, VLA, WAM in an indiscriminate fashion, as is becoming common in robotics, we are not helping clarity in communication. Let's keep the h

researchyann-lecun--x
23 Aug 2026
Local Ai

“The All Spark” Cluster: Upgrading from 16 - 36 DGX Sparks

DGX agent

Earlier this year I posted about building what at the time I believe was the first 16x DGX Spark Cluster. I’m now adding 20 more Sparks to the cluster in my homelab server rack, giving me 4.6TB of uni

local-air-localllama
23 Aug 2026
Model Releases

the difference in productivity between 0-1 mode vs. working within a big team at a large co seems universally true even pre-chatgpt. it's re…

DGX agent

the difference in productivity between 0-1 mode vs. working within a big team at a large co seems universally true even pre-chatgpt. it's really easy to be 'productive' when (a) you're the only person

model-releasesjerry-liu--x
23 Aug 2026
← Previous
1…768769770771772…1369
Next →