AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,675 results
Model Releases

Decoupled Descent: Enforcing Exact Train-Test Error Tracking Via AMP Onsager Corrections [R]

DGX agent

Link: https://arxiv.org/pdf/2604.27883 Hi, Most of use are familiar with the headache of training a neural network using gradient descent where the training error may go to zero but the test error may

model-releasesr-machinelearning
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

DeepSeek-V4-Flash-0731 (284B MoE) at 75 tok/s on 2× DGX Spark — full recipe, 11 gotchas, reboot-proof cluster, Codex CLI integration

DGX agent

Spent two nights getting deepseek-ai/DeepSeek-V4-Flash-0731 (284B MoE, 13B active, native FP4/FP8, 1M context) running production-grade on two DGX Sparks connected by one QSFP DAC cable. Everything —

model-releasesr-localllama
11 Aug 2026
Model Releases

Full-bandwidth transformer

DGX agent

arXiv:2608.08888v1 Announce Type: new Abstract: Autoregressive transformers compute along two axes: horizontally across generated tokens, and vertically through model depth. Dense attention gives each

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

GeoRoute: Geometry-Aware Hybrid Inference for Traffic Future-Frame Prediction

DGX agent

arXiv:2608.09493v1 Announce Type: new Abstract: Long-horizon future-frame prediction is important for autonomous driving, traffic surveillance, and intelligent transportation systems, yet remains chal

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Harmful Content Is Not Enough: Continuation Framing Moderates In-Context Emergent Misalignment

DGX agent

arXiv:2608.08212v1 Announce Type: new Abstract: In-context learning (ICL) can induce emergent misalignment (EM), where narrow misaligned examples alter answers to unrelated questions. Existing prompts

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Hidden Language Consistency Phenomena in Reasoning LLMs

DGX agent

arXiv:2608.08447v1 Announce Type: cross Abstract: Multilingual reasoning models are commonly evaluated by whether they arrive at the correct answer, but not by whether they preserve the intended langu

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

How Much Does It Cost to Answer My Question? Benchmarking Cloud VLM-based VQA Systems

DGX agent

arXiv:2608.07861v1 Announce Type: new Abstract: Vision-language models (VLMs) are becoming a practical backend for mobile visual question answering (VQA) systems, enabling smartphones and smart glasse

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

I ran Muse Glimmer @ 1M context - All tests passed.

DGX agent

Heeeey all! I just completed some fun tests with Muse Glimmer, I thought I'd let you know. In fact, the summary below was written by Muse itself! I ran a 2× DGX Spark cluster and got Meta's day-old Mu

model-releasesr-localllama
11 Aug 2026
Model Releases

ICM Out! Better Tournament Strategy from Computed Continuations, vs. Solvers and LLMs

DGX agent

arXiv:2608.09586v1 Announce Type: new Abstract: The Independent Chip Model (ICM) converts tournament chips into reference prize equity, and policies are routinely constructed against those values. Bec

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

IndexTTS 2.5 Technical Report

DGX agent

arXiv:2601.03888v4 Announce Type: replace-cross Abstract: In prior work, we introduced IndexTTS 2, a zero-shot neural text-to-speech foundation model comprising two core components: a transformer-base

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Learning to Triage Vulnerability Reports from Program Analysis: An Empirical Study in Node.js

DGX agent

arXiv:2510.20739v2 Announce Type: replace-cross Abstract: Program analysis tools often produce large volumes of candidate vulnerability reports that require costly manual review, creating a practical

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

LexKairos: Benchmarking Legal Temporal Capabilities in LLMs

DGX agent

arXiv:2608.09106v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated strong performance across a wide range of legal tasks. In legal practice, time is a critical concept that

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Looker’s semantic layer governs Gemini Enterprise data for user trust

DGX agent

For organizations deploying AI agents at scale, there’s often a critical divide between structured and unstructured data. While large language models (LLMs) excel at parsing text documents, emails, an

model-releasesgoogle-cloud-ai
11 Aug 2026
Model Releases

LoRSA: Toward Generalizable Parameter-Efficient Fine-Tuning for Biomedical Downstream Tasks

DGX agent

arXiv:2608.07749v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning enables the adaptation of vision foundation models to biomedical tasks under limited computational resources, but a si

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Math-Vision Diagrams: A Comprehensive Benchmark for Evaluating LLM Mathematical Diagram Generation Capabilities

DGX agent

arXiv:2608.08964v1 Announce Type: new Abstract: The generation of mathematically precise diagrams from tex- tual prompts has emerged as a critical yet underexplored capability of Large Language Models

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

MemeMind: Reference-Guided Trace Construction for Offline Context Optimization

DGX agent

arXiv:2608.09316v1 Announce Type: new Abstract: Offline context optimization improves an agent by revising its instructions and examples while keeping the model frozen. This approach learns from rollo

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

MMArch: Benchmarking Multimodal Reasoning Grounded in Architectural Evidence

DGX agent

arXiv:2608.09281v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) perform strongly on engineering imagery, yet existing benchmarks mostly test drawing recognition, information e

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Persistent Semantic Entities in Tool-Augmented LLM Systems

DGX agent

arXiv:2608.07952v1 Announce Type: cross Abstract: Tool-augmented LLM agents can harbor implicit state that persists across sessions, activates through events, and propagates across agent boundaries---

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Same Question, Different Answer? Measuring and Mitigating Prompt Privilege for Equitable AI Access

DGX agent

arXiv:2608.08942v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into healthcare, education, public services, and everyday decision making. They should provide

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Shattered Compositionality: Counterintuitive Learning Dynamics of Transformers for Arithmetic

DGX agent

arXiv:2601.22510v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often achieve strong benchmark accuracy yet remain brittle under small distribution shifts. While recent mechanis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests

DGX agent

arXiv:2608.08640v1 Announce Type: new Abstract: Large language model agents increasingly rely on reusable skills to extend their capabilities beyond parametric knowl- edge. However, retrieving the app

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation

DGX agent

arXiv:2601.09974v2 Announce Type: replace Abstract: Personalizing Large Language Models typically relies on static retrieval or one-time adaptation, assuming user preferences remain invariant over tim

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Targeted Counterfactual Fingerprinting for Black-Box LLM Ownership Verification

DGX agent

arXiv:2608.08195v1 Announce Type: cross Abstract: Large language models (LLMs) are high-value assets that can be derived through redeployment, fine-tuning, quantization, or further alignment. Because

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

The Authority Expectancy Effect in Multi-User Conflict

DGX agent

arXiv:2608.08026v1 Announce Type: new Abstract: We investigate how social authority (SA) signals interact with severity-based prioritization in large language models, operationalizing each axis as a m

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

The Collaboration Gap: Exploration and Benchmarking of Open-World Agentic Cooperation

DGX agent

arXiv:2511.02687v2 Announce Type: replace Abstract: The trajectory of AI development suggests that we will increasingly rely on agent-based systems powered by language models, composed of independentl

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Understanding Calibration and Truncation Error Propagation in Training-Free Low-Rank Compression for LLMs

DGX agent

arXiv:2608.08506v1 Announce Type: new Abstract: Training-free low-rank compression frameworks have been gaining prominence for LLM compression given their effectiveness in reducing model parameter cou

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

v0.32.8

DGX agent

Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such

model-releasesollama-releases
11 Aug 2026
Research

When Confidence Fails: Overconfidence in LLMs under Uncertainty and Missing Clinical Information

DGX agent

arXiv:2608.09080v1 Announce Type: cross Abstract: Large Language Models (LLMs) have achieved strong performance in medical question answering and clinical reasoning tasks. However, their reliability u

researcharxiv-cs-ai
11 Aug 2026
Model Releases

b10338

DGX agent

model-saver : fix expert shared/chunk FFN length key clobber (#26693) The saver called add_kv with LLM_KV_EXPERT_SHARED_FEED_FORWARD_LENGTH twice, the second time passing n_ff_chexp. gguf_set_val_u32

model-releasesllama-cpp-releases
10 Aug 2026
Model Releases

Best open-source harness like Claude Code?

DGX agent

Avid claude code user here looking to do equivalent things with local models. Just want to plug in something like Qwen and have the interface be 1:1 with claude code. Any suggestion? submitted by /u/N

model-releasesr-localllama
10 Aug 2026
Safety

CASA: Classification Augmented with Safety Attention for Robust Multimodal Alignment

DGX agent

arXiv:2604.00310v2 Announce Type: replace-cross Abstract: Multimodal large-language models (MLLMs) often experience degraded safety alignment when harmful queries exploit cross-modal interactions. Mod

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Corma launches with $60M in funding for defensive cybersecurity AI

DGX agent

Defensive cybersecurity startup Corma Labs Ltd. today announced it has raised 60 million in seed funding to build a foundation model purpose-built for security defense. Founded in 2025, Corma runs off

model-releasessiliconangle
10 Aug 2026
Model Releases

Cryptanalytic Extraction of Isolated Bias-Free GLU Feed-Forward Blocks by Antipodal Separation

DGX agent

arXiv:2608.06631v1 Announce Type: cross Abstract: Cryptanalytic extraction has been demonstrated for ReLU networks, for networks using componentwise activations such as GELU or SiLU, and for a Transfo

model-releasesarxiv-cs-ai
10 Aug 2026
Research

How Long Reasoning Chains Influence LLMs' Judgment of Answer Factuality

DGX agent

arXiv:2604.06756v2 Announce Type: replace Abstract: Large language models (LLMs) has been widely adopted as a scalable surrogate for human evaluation, yet such judges remain imperfect and susceptible

researcharxiv-cs-cl
10 Aug 2026
Safety

Let's Unlearn Stereotypes Before Decision-Making: Assessing the Impact of Intrinsic Bias Mitigation on Downstream Fairness in LLMs

DGX agent

arXiv:2509.16462v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly used in high-stakes decision-making systems, where biased predictions can reinforce social and economi

safetyarxiv-cs-cl
10 Aug 2026
Research

Natural Language Processing Psychometrics

DGX agent

arXiv:2608.07316v1 Announce Type: cross Abstract: Natural Language Processing (NLP) models predicting mental health outcomes rarely specify what they measure: contextual knowledge, emotional content,

researcharxiv-cs-ai
10 Aug 2026
Model Releases

Semantic Adapter Routing with Fine-Tuning Task Embeddings

DGX agent

arXiv:2606.19079v2 Announce Type: replace Abstract: Parameter-efficient fine-tuning (PEFT) has led to model ecosystems in which a single backbone is paired with many task-specialized adapters. Given s

model-releasesarxiv-cs-ai
10 Aug 2026
Tutorials

Semi Edge Inference Idea [D]

DGX agent

Today the most important factor in AI is cost. My idea is to split ML models inference (closed ones, proprietary) across server and edge computing on clients, and I would like to hear what do you thin

tutorialsr-machinelearning
10 Aug 2026
Model Releases

SLED: Scalable Location Encoding via Distillation

DGX agent

arXiv:2608.06612v1 Announce Type: cross Abstract: The plethora of readily available geospatial data offers exciting opportunities to learn high quality representations of the planet, but the sheer siz

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

Stable Curves, Unstable Items: Item-Level Scaling Heterogeneity in Video LLMs

DGX agent

arXiv:2608.07014v1 Announce Type: new Abstract: Aggregate scaling curves suggest that Video LLMs improve smoothly or saturate as visual budgets grow. We show that this view can conceal large, opposing

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

Stockmark-Nemotron-3-Nano-Omni-JapanDocReader: Structured Document Parsing via Capability Injection and Forgetting Control

DGX agent

arXiv:2608.06758v1 Announce Type: new Abstract: We present Stockmark-Nemotron-3-Nano-Omni-JapanDocReader, a Japanese document understanding model built from Nemotron-3-Nano-Omni-30B-A3B-Reasoning-BF16

model-releasesarxiv-cs-cl
10 Aug 2026
Model Releases

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

DGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

model-releasesdair-ai--x
10 Aug 2026
Model Releases

The Sparsity Whisperer

DGX agent

arXiv:2608.06630v1 Announce Type: new Abstract: Pruning reduces the inference cost of large language models, but existing criteria primarily preserve large activations or reconstruct layer outputs. We

model-releasesarxiv-cs-lg
10 Aug 2026
Model Releases

UAV3DCrop: Benchmarking 3D Reconstruction in Repeated Multi-Angle UAV Crop Surveys

DGX agent

arXiv:2608.06404v1 Announce Type: new Abstract: Accurate 3D crop monitoring underpins data-driven precision agriculture by enabling field-scale analysis of plant structure, growth dynamics, and manage

model-releasesarxiv-cs-cv
10 Aug 2026
Model Releases

We've used GPT-5.6-Cyber extensively in real-world vulnerability research, including work that uncovered previously unknown vulnerabilities …

DGX agent

OpenAI announced the release of GPT‑5.6‑Cyber as part of its Cybersecurity Initiative, “Daybreak.” The model is aimed at advanced, authorized security research and testing, helping trusted defenders d

model-releasesopenai--x
10 Aug 2026
Model Releases

Zero Gap Is Not Restoration: Stratified Per-Question Probability Evaluation and Step-wise Mitigation of Benchmark Contamination

DGX agent

arXiv:2608.07341v1 Announce Type: cross Abstract: Test data from public benchmarks inevitably leaks into pretraining corpora, inflating evaluation scores once memorized. extbf{Contamination mitigation

model-releasesarxiv-cs-ai
10 Aug 2026
Model Releases

[2606.05682] Beyond Output Matching: Preserving Internal Geometry in NVFP4 LLM Distillation

DGX agent

Demand for low-precision inference, including NVFP4-based approaches, has grown as large language models are increasingly deployed in latency and cost constrained production environments. Quantization

model-releasesr-localllama
9 Aug 2026
Model Releases

The Gemma team will host a special event on August 20

DGX agent

Tweet by u/hackerllama Could be copium, but I would love to see Gemma 4.1 there with unified audio input for all model sizes perhaps even up to 120B, much improved tool calling (even with the latest t

model-releasesr-localllama
9 Aug 2026
← Previous
1…394395396397398…1369
Next →