AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,612 results
Model Releases

SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents

DGX agent

arXiv:2605.21384v1 Announce Type: cross Abstract: As long-horizon coding agents produce more code than any developer can review, oversight collapses onto a single surface: the automated test suite. Re

model-releasesarxiv-cs-cl
21 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining

DGX agent

arXiv:2605.20296v1 Announce Type: new Abstract: Fine-tuning a language model for a target task routinely degrades capabilities the training data never explicitly threatened. We study this phenomenon,

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Spent Tuesday watching @Google I/O thinking about a question. If you build on @Android, what just changed for you? Short answer: your screen…

DGX agent

Spent Tuesday watching @Google I/O thinking about a question. If you build on @Android, what just changed for you? Short answer: your screen belongs to Gemini now. After this week, Google owns the AI,

model-releasesdiv-garg--x
21 May 2026
Model Releases

Spotify Labs launches Studio, a NotebookLM-like desktop app to generate private, AI-powered podcasts, in research preview across 20+ markets (Ivan Mehta/TechCrunch)

DGX agent

Ivan Mehta / TechCrunch: Spotify Labs launches Studio, a NotebookLM-like desktop app to generate private, AI-powered podcasts, in research preview across 20+ markets — One of the common features for c

model-releasestechmeme
21 May 2026
Model Releases

Spotify says it has 1M+ subscriptions to Audiobook+, which is on track for $100M in ARR, and unveils an ElevenLabs-powered audiobook self-publishing tool (Ivan Mehta/TechCrunch)

DGX agent

Ivan Mehta / TechCrunch: Spotify says it has 1M+ subscriptions to Audiobook+, which is on track for $100M in ARR, and unveils an ElevenLabs-powered audiobook self-publishing tool — Alongside tools for

model-releasestechmeme
21 May 2026
Model Releases

STAR-IOD: Scale-decoupled Topology Alignment with Pseudo-label Refinement for Remote Sensing Incremental Object Detection

DGX agent

arXiv:2605.20738v1 Announce Type: new Abstract: Remote sensing imagery typically arrives in the form of continuous data streams. Traditional detectors often forget previously learned categories when l

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

SymbolicLight V1: Spike-Gated Dual-Path Language Modeling with High Activation Sparsity and Sub-Billion-Scale Pre-Training Evidence

DGX agent

arXiv:2605.21333v1 Announce Type: new Abstract: Natively trained spiking language models struggle to combine Transformer-like language quality, stable multi-domain pre-training, and high activation sp

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

TabPFN Extensions for Interpretable Geotechnical Modelling

DGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design

DGX agent

arXiv:2605.20731v1 Announce Type: new Abstract: Text-to-image models produce graphic design at production scale, but their supervision comes from photo-style preference data with a single overall verd

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

TempGlitch: Evaluating Vision-Language Models for Temporal Glitch Detection in Gameplay Videos

DGX agent

arXiv:2605.21443v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly being explored for video game quality assurance, especially gameplay glitch detection. Most existing eval

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

DGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Text Analytics Evaluation Framework: A Case Study on LLMs and Social Media

DGX agent

arXiv:2605.21338v1 Announce Type: new Abstract: LLMs have demonstrated exceptional proficiency in a wide range of NLP tasks. However, a notable gap remains in practical data analysis scenarios, partic

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

TextSculptor: Training and Benchmarking Scene Text Editing

DGX agent

arXiv:2605.21090v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and diffusion-based generative models have substantially improved prompt-driven image editin

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets

DGX agent

arXiv:2605.20279v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

The top announcements for startups from Google I/O ‘26

DGX agent

Many of the world’s fastest-growing AI startups are choosing to build their future — and the world’s — on Google Cloud because of our complete and open AI stack. We embed AI into every layer of our ar

model-releasesgoogle-cloud-ai
21 May 2026
Model Releases

The Visual Iconicity Challenge: Evaluating Vision-Language Models on Sign Language Form-Meaning Mapping

DGX agent

arXiv:2510.08482v3 Announce Type: replace-cross Abstract: Iconicity, the resemblance between linguistic form and meaning, is pervasive in signed languages, offering a natural testbed for visual ground

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

The Yes-Man Syndrome: Benchmarking Abstention in Embodied Robotic Agents

DGX agent

arXiv:2605.20544v1 Announce Type: cross Abstract: Vision-language models (VLMs) are used as high-level planners for embodied agents, translating natural language instructions and visual observations i

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Theoretical guidelines for annealed Langevin dynamics in compositional simulation-based inference

DGX agent

arXiv:2605.21253v1 Announce Type: cross Abstract: Compositional score-based approaches to simulation-based inference (SBI) approximate the posterior over a shared parameter given n independent observa

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

THEval. Evaluation Framework for Talking Head Video Generation

DGX agent

arXiv:2511.04520v4 Announce Type: replace Abstract: Video generation has achieved remarkable progress, with generated videos increasingly resembling real ones. However, the rapid advance in generation

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Time-Dependent PDE-Constrained Optimization via Weak-Form Latent Dynamics

DGX agent

arXiv:2605.20639v1 Announce Type: cross Abstract: Optimization problems constrained by high-dimensional, time-dependent partial differential equations require repeated forward and sensitivity solves,

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

TimeSRL: Generalizable Time-Series Behavioral Modeling via Semantic RL-Tuned LLMs -- A Case Study in Mental Health

DGX agent

arXiv:2605.21295v1 Announce Type: new Abstract: Longitudinal passive sensing enables continuous health prediction, yet models often fail under cross-dataset distribution shifts. Traditional ML overfit

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Tippett-minimum Fusion of Representation-space Diffusion Models for Multi-Encoder Out-of-Distribution Detection

DGX agent

arXiv:2605.20502v1 Announce Type: cross Abstract: We address out-of-distribution (OOD) detection across the full spectrum of distribution shifts -- global domain changes, semantic divergence, texture

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million par…

DGX agent

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million parameters and runs nearly 2,000 tokens per second prefill on a

model-releasesswyx--x
21 May 2026
Model Releases

To Select or not to Select, that is the Question: Distilling Robot Skill Prediction into a Small Ensemble

DGX agent

arXiv:2605.21242v1 Announce Type: new Abstract: As robot fleets become more heterogeneous, including humanoids, rovers, quadrupeds, and drones, selecting the right robot for a task becomes a core syst

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefi…

DGX agent

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefit one at a time inside a 1.7B byte-level pretraining pipelin

model-releasesnous-research--x
21 May 2026
Model Releases

Token Maxxing, April 2026-May 2026, RIP 🪦

DGX agent

Token Maxxing, April 2026-May 2026, RIP 🪦 🦔Microsoft canceled its internal Claude Code licenses this week after token-based billing made the cost untenable, even for a company with effectively infinit

model-releasesgary-marcus--x
21 May 2026
Model Releases

Towards UAV Detection in the Real World: A New Multispectral Dataset UAVNet-MS and a New Method

DGX agent

arXiv:2605.20963v1 Announce Type: new Abstract: The proliferation of unmanned aerial vehicles (UAVs) has created urgent demand for precise UAV monitoring. Existing RGB-based systems rely on spatial cu

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Toxic Subword Pruning for Dialogue Response Generation on Large Language Models

DGX agent

arXiv:2410.04155v2 Announce Type: replace Abstract: How to defend large language models (LLMs) from generating toxic content is an important research area. Yet, most research focused on various model

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Training distribution determines the ceiling of drug-blind cancer sensitivity prediction

DGX agent

arXiv:2605.20885v1 Announce Type: new Abstract: Precision oncology requires predicting which drugs will suppress a specific tumor from its molecular profile, but drug-blind sensitivity prediction has

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Training Language Agents to Learn from Experience

DGX agent

arXiv:2605.20477v1 Announce Type: cross Abstract: Language agents can adapt from experience in interactive environments, but current reflection-based methods can only self-correct within a single task

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

TRAM: Test-Time Risk Adaptation with Mixture of Agents

DGX agent

arXiv:2408.08812v2 Announce Type: replace Abstract: Deployed reinforcement learning agents often face safety requirements that are specified only after training, such as new hazard maps, revised risk

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Try Composer 2.5

DGX agent

Try Composer 2.5 New CursorBench results just dropped. Two big takeaways. Composer 2.5 is way better than most people think. 63.2% score at $0.55 per task. Nearly matching Opus 4.7 Max and GPT 5.5 Ext

model-releaseselon-musk--x
21 May 2026
Model Releases

Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning

DGX agent

arXiv:2605.20803v1 Announce Type: new Abstract: Continual learning (CL) aims to train models sequentially on multiple tasks while mitigating catastrophic forgetting of previously learned knowledge. Re

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models

DGX agent

arXiv:2605.20202v1 Announce Type: new Abstract: I study whether emotionally framed evaluation follow-ups change both the behavior and the calm-relative internal representations of small, locally deplo

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Understanding and Improving Communication Performance in Multi-node LLM Inference

DGX agent

arXiv:2511.09557v4 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in size, distributed inference has become increasingly important. Model-parallel strategies m

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs

DGX agent

arXiv:2505.19075v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable general capabilities, but enhancing skills such as reasoning often demands substanti

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

VDFP: Video Deflickering with Flicker-banding Priors

DGX agent

arXiv:2605.21079v1 Announce Type: new Abstract: Capturing digital screens with smartphones frequently induces severe banding due to hardware synchronization mismatches. Existing video restoration meth

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Verifiable Provenance and Watermarking for Generative AI: An Evidentiary Framework for International Operational Law and Domestic Courts

DGX agent

arXiv:2605.21002v1 Announce Type: cross Abstract: Generative artificial intelligence now synthesizes photorealistic imagery, audio, and video at a cost that defeats traditional forensic intuition. The

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

VersusQ: Pairwise Margin Reasoning for Generalizable Video Quality Assessment

DGX agent

arXiv:2605.21130v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have shown promise for video quality assessment, but most methods still predict an absolute score for each video. Such po

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Vision Transformers and Convolutional Neural Networks for Land Use Scene Classification

DGX agent

arXiv:2605.21268v1 Announce Type: new Abstract: Land Use Scene Classification (LUSC) from remote sensing imagery plays a critical role in environmental monitoring, urban planning, and sustainable reso

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

VISTA: Technical Report for the Ego4D Short-Term Object Interaction Anticipation at EgoVis 2026

DGX agent

arXiv:2605.20901v1 Announce Type: new Abstract: We propose VISTA, a V-JEPA Integrated StillFast Temporal Anticipator for the Ego4D Short-Term Object Interaction Anticipation (STA) Challenge at EgoVis

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

VISTAQA: Benchmarking Joint Visual Question Answering and Pixel-Level Evidence

DGX agent

arXiv:2605.20676v1 Announce Type: new Abstract: Establishing a clear link between model predictions and the visual evidence that supports them is critical for transparency and reliability in multimoda

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models

DGX agent

arXiv:2605.20774v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong promise for general-purpose robotic manipulation, but their real-world evaluation remains limited

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

VSCD: Video-based Scene Change Detection in Unaligned Scenes

DGX agent

arXiv:2605.20821v1 Announce Type: new Abstract: Detecting what has changed in an environment is essential for long-term autonomy, yet most change detection settings assume fixed viewpoints, mild misal

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

WaveGraphNet: Physics-Consistent Guided-Wave Damage Localization through Coupled Inverse-Forward Graph Learning

DGX agent

arXiv:2605.20311v1 Announce Type: new Abstract: Guided-wave structural health monitoring enables damage localization in composite plates using sparse networks of bonded piezoelectric transducers. Howe

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

WCXB: A Multi-Type Web Content Extraction Benchmark

DGX agent

arXiv:2605.21097v1 Announce Type: new Abstract: Web content extraction - isolating a page's main content from surrounding boilerplate - is a prerequisite for search indexing, retrieval-augmented gener

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics

DGX agent

arXiv:2605.20441v1 Announce Type: new Abstract: Transformers trained on modular arithmetic exhibit sharp transitions between memorization, generalization, and collapse. We show that weight decay acts

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

We’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental risks

DGX agent

Google DeepMind has launched 'AI for the Planet,' a three-month accelerator program in Asia Pacific focused on leveraging advanced AI to combat environmental challenges like climate change, biodiversi

model-releasesgoogle-deepmind
21 May 2026
← Previous
1…282283284285286…472
Next →