AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,280 results
Model Releases

Rethinking Attention Locality in Spiking Transformers

DGX agent

arXiv:2608.08541v1 Announce Type: new Abstract: Spiking Transformers provide a promising paradigm for efficient visual processing with spike-driven computation, yet their Softmax-free Spiking Self-Att

model-releasesarxiv-cs-cv
11 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Rethinking Medical Landmark Localization with Prototype Learning-based Progressive Offset Correction

DGX agent

arXiv:2608.09182v1 Announce Type: cross Abstract: Accurate landmark localization in medical images is a fundamental step for quantitative clinical measurement and downstream analysis. Existing localiz

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Rethinking Self-Evolving Agents: Do We Still Need Prescribed Optimization Pipelines?

DGX agent

arXiv:2608.09629v1 Announce Type: new Abstract: Self-evolving agents are usually built around prescribed optimization pipelines: the framework decides how to gather evidence, revise a persistent artif

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RippleKV: Cross-Layer KV Cache Allocation via Perturbation Propagation

DGX agent

arXiv:2608.08684v1 Announce Type: cross Abstract: Long-context LLM inference is bottlenecked by KV cache memory, yet distributing a limited cache budget across layers remains challenging. Existing met

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RISE-RL: Rubric-Informed Selective Exploration for Open-Ended Reinforcement Learning

DGX agent

arXiv:2608.09123v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) for open-ended tasks is challenging because responses must satisfy multidimensional criteria without following a s

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RotaryQuant: Fitting 120B MoE Models on Consumer Hardware via Fused Compressed-Space Attention

DGX agent

arXiv:2608.08081v1 Announce Type: cross Abstract: Large mixture-of-experts (MoE) language models with 26--120 billion parameters exceed the memory capacity of consumer devices through three simultaneo

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

RouteGuard: Certifying Routing Gain in LLM Multi-Agent Systems When Complementarity Is Not Enough

DGX agent

arXiv:2608.07583v1 Announce Type: cross Abstract: Multi-agent LLM systems route among model-backed advisors, yet a deployer rarely knows before shipping whether routing will help at all. Prevailing ro

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Router Sensitivity Under Lightweight Fine-Tuning Identifies Prunable Experts in Mixture-of-Experts Models

DGX agent

arXiv:2608.07890v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models decouple total parameters from per-token compute, but deployment still requires storing every expert. Recent theory sh

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SafeSceneReason: A Multimodal Reasoning Benchmark Connecting Industrial Hazards with Accident Knowledge

DGX agent

arXiv:2608.09230v1 Announce Type: new Abstract: Industrial-safety understanding requires more than detecting workers, equipment, and personal protective equipment. Models must also assess compliance,

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SAGE: SLO-Aware Adaptive Retrieval for Production RAG Systems

DGX agent

arXiv:2608.08237v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) systems in production operate under strict service level objectives (SLOs) on tail latency and infrastructure cos

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

SAIN: Structure-Aware Interactive Navigation with Active Dialogue Grounding for Mobile Robot

DGX agent

arXiv:2608.09196v1 Announce Type: new Abstract: Most existing vision-language navigation tasks assume that instructions are complete and unambiguous. However, real-world robots often encounter natural

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Same Question, Different Answer? Measuring and Mitigating Prompt Privilege for Equitable AI Access

DGX agent

arXiv:2608.08942v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into healthcare, education, public services, and everyday decision making. They should provide

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Sci-VBench: Evaluating Knowledge- and Reasoning-Intensive Video Generation in Science Domains

DGX agent

arXiv:2608.09873v1 Announce Type: cross Abstract: We introduce Sci-VBench, a comprehensive benchmark for evaluating knowledge- and reasoning-intensive video generation across scientific domains. It co

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SciTaRC: A Plan-Annotated Scientific Tabular QA Benchmark for Language Reasoning and Complex Computation

DGX agent

arXiv:2603.08910v2 Announce Type: replace Abstract: We introduce SciTaRC, an expert-authored benchmark for question answering over scientific tables that targets composite, multi-step reasoning. To en

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

SCTD 3.0: Sonar Common Target Detection in the Wild - A Large-Scale, Multi-Scene Dataset from Real Marine Surveys

DGX agent

arXiv:2608.08106v1 Announce Type: new Abstract: Synthetic Aperture Sonar (SAS) is core for wide-area detection of small underwater targets. However, large-scale, high-quality SAS datasets are scarce,

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

SDDBMs: Soft Denoising Diffusion Bridge Models

DGX agent

arXiv:2608.08594v1 Announce Type: new Abstract: Diffusion bridge models leverage Doob's (h)-transform to construct stochastic transports between arbitrary endpoint distributions, and have shown strong

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sekai2: From World Exploration to Interactive World Modeling

DGX agent

arXiv:2608.09449v1 Announce Type: new Abstract: Video world models must capture how scenes evolve over time and across viewpoints. Training them for long-horizon generation and camera control therefor

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

SeqLoc: Beyond the Single Frame for Cross-View Geo-Localization in Feature-Sparse Scenes

DGX agent

arXiv:2608.07835v1 Announce Type: new Abstract: Cross-View Geo-Localization (CVGL) with OpenStreetMap (OSM) performs well in structure-rich urban environments but collapses in feature-sparse scenes su

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Shape Mutating Expert Compression:LorExperts and BTExperts

DGX agent

arXiv:2608.07814v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) language models deliver high capacity at low per-token compute, but deploying them cheaply requires compressing their many ex

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Shattered Compositionality: Counterintuitive Learning Dynamics of Transformers for Arithmetic

DGX agent

arXiv:2601.22510v2 Announce Type: replace-cross Abstract: Large language models (LLMs) often achieve strong benchmark accuracy yet remain brittle under small distribution shifts. While recent mechanis

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SHE: Trajectory-driven Safety Harness Evolution for LLM Agents

DGX agent

arXiv:2608.09885v1 Announce Type: new Abstract: The safety of large language model (LLM) agents depends not only on model weights but also on the agent harness that manages context, memory, tools, per

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SI-Edit: Toward Sketch-Instruction Guided Local Image Editing with Pixel-Level Precision

DGX agent

arXiv:2608.09097v1 Announce Type: new Abstract: Despite rapid advances in generative models, achieving pixel-level precision in sketch-based image editing remains a persistent challenge, particularly

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

SIMMER: Benchmarking Latent Failures in LLM Executable Planning with a World Model

DGX agent

arXiv:2606.14574v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as planners for autonomous agents in household environments. While existing benchmarks

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SkillConsist: Detecting Inconsistencies in Agent Skills via Bidirectional Graph Alignment

DGX agent

arXiv:2608.07639v1 Announce Type: cross Abstract: Agent Skills provide reusable capabilities to LLM agents. Agent Skill inconsistencies can expose undisclosed dangerous behavior or cause wrong Skill s

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests

DGX agent

arXiv:2608.08640v1 Announce Type: new Abstract: Large language model agents increasingly rely on reusable skills to extend their capabilities beyond parametric knowl- edge. However, retrieving the app

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SkillSentry: Reliable Skill Execution for LLM Agents via Runtime Assurance

DGX agent

arXiv:2608.09253v1 Announce Type: new Abstract: LLM agents are increasingly equipped with skills to perform complex tasks through multi-step reasoning and tool use. Although skills provide reusable pr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SLIM-0.5B: Learning Action-Grounded Predictive Latents for Robot Manipulation

DGX agent

arXiv:2608.09771v1 Announce Type: new Abstract: Vision-language-action policies rely on large multimodal backbones to jointly perform perception, language conditioning, and action generation at every

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Smart Compaction: Predicting Compaction Utility from Lakehouse Table Metadata

DGX agent

arXiv:2608.08639v1 Announce Type: new Abstract: Open lakehouse table formats accumulate small data files over time, which degrades query performance. Deciding when compaction is worthwhile remains thr

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Social Gym and SPaRTan: Benchmarking and Improving LLM Social Reasoning via Multi-Agent Game Tournaments

DGX agent

arXiv:2608.09128v1 Announce Type: cross Abstract: LLM agents are increasingly deployed in multi-agent social settings where they must cooperate, negotiate, and adapt to other agents. Measuring and imp

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SodaMem: Evidence-Grounded Temporal Graph Memory for LLM Agents

DGX agent

arXiv:2608.08055v1 Announce Type: new Abstract: Large language model (LLM) agents that assist users over weeks of conversation must remember what is currently true, not merely what was once said. Flat

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Sources: Nvidia is developing a Nemotron 4 model with 1T+ parameters, up from Nemotron 3 Ultra's 550B parameters but smaller than leading Chinese open models (The Information)

DGX agent

The Information: Sources: Nvidia is developing a Nemotron 4 model with 1T+ parameters, up from Nemotron 3 Ultra's 550B parameters but smaller than leading Chinese open models — Nvidia is doubling down

model-releasestechmeme
11 Aug 2026
Model Releases

Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi Agents

DGX agent

arXiv:2601.18077v3 Announce Type: replace Abstract: Cooperative reasoning under incomplete information remains challenging for both humans and multi-agent systems. The card game Hanabi embodies this c

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

Sparse corruption in low-rank matrix inference: the PCA benchmark

DGX agent

arXiv:2511.11927v2 Announce Type: replace-cross Abstract: Principal Component Analysis (PCA) is a standard tool for extracting a low-rank signal from noisy observations. It is known that applying PCA

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

SPECTRA: Pushing the KV Cache Beyond the 2-Bit Cliff via Spectral Transform Coding

DGX agent

arXiv:2608.07915v1 Announce Type: new Abstract: Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns. Th

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Spectral Outliers Reveal Dominant Learned Structure in Transformer Attention

DGX agent

arXiv:2608.07921v1 Announce Type: cross Abstract: We apply Marchenko-Pastur (MP) random matrix theory to pre-trained attention weights in order to separate each projection matrix into a random-like bu

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SpikeWorld: Fast-State Adaptation for Frozen Spiking World Models

DGX agent

arXiv:2608.07712v1 Announce Type: cross Abstract: A predictive model receives a self-supervised signal whenever the consequence of an action is observed. Using that signal after deployment is difficul

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation

DGX agent

arXiv:2601.09974v2 Announce Type: replace Abstract: Personalizing Large Language Models typically relies on static retrieval or one-time adaptation, assuming user preferences remain invariant over tim

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

DGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

Staying True to the Origin: Continuous Image Stylization with Smooth Transitions

DGX agent

arXiv:2608.08125v1 Announce Type: new Abstract: Recent advances in generative models have achieved remarkable performance in text- and image-conditioned editing. However, preserving the content of a g

model-releasesarxiv-cs-cv
11 Aug 2026
Model Releases

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

Stealing Reasoning Traces from Proprietary LLM APIs A vanity domain name (stolen-thoughts.com) for a neat paper: Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients that

model-releasessimon-willison
11 Aug 2026
Model Releases

Subjective Multi-Bias Detection with Large Language Models

DGX agent

arXiv:2608.09126v1 Announce Type: new Abstract: In this project, we delved into the pervasive challenge of bias detection within the text content. More specifically, our focus lies on the identificati

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

SUM-AgriVLN: Spatial Understanding Memory for Agricultural Vision-and-Language Navigation

DGX agent

arXiv:2510.14357v2 Announce Type: replace-cross Abstract: Agricultural robots are emerging as powerful assistants across a wide range of agricultural tasks, nevertheless, they are still heavily relyin

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SuperCoder: Assembly Program Superoptimization with Large Language Models

DGX agent

arXiv:2505.11480v4 Announce Type: replace-cross Abstract: Superoptimization is the task of transforming a program into a faster one, and ideally the very fastest possible one, while preserving its inp

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SuperLocalMemory 4.0: The Governed Memory Operating System for AI Agents

DGX agent

arXiv:2608.08253v1 Announce Type: new Abstract: AI agents are becoming shared infrastructure, yet durable memory is commonly assembled from separate retrieval, governance, and operational components.

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SurakshaEval: An Indic Safety Benchmark for Multilingual LLMs

DGX agent

arXiv:2608.07862v1 Announce Type: new Abstract: Existing safety evaluation datasets for large language models (LLMs) predominantly focus on English and Western contexts, often overlooking the linguist

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

SurgWMBench: A Vision-Based Benchmark for World-Modeling Surgical Instrument Motion Planning

DGX agent

arXiv:2608.08070v1 Announce Type: new Abstract: Reliable surgical planning requires models that move beyond recognizing the current surgical step or imitating expert demonstrations, and instead antici

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

SurveyReview: A Reviewer-Aligned Benchmark for Survey Evaluators

DGX agent

arXiv:2608.07641v1 Announce Type: new Abstract: The rapid advancement of large language models has transformed survey writing from a months-long manual effort into an automated process. As generation

model-releasesarxiv-cs-cl
11 Aug 2026
Model Releases

SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring

DGX agent

arXiv:2608.09802v1 Announce Type: new Abstract: As AI coding agents take on increasingly complex, long-horizon software engineering tasks, existing benchmarks are rapidly saturating and their evaluati

model-releasesarxiv-cs-cl
11 Aug 2026
← Previous
1…1617181920…465
Next →