AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,585 results
21 May 2026

Spotify Labs launches Studio, a NotebookLM-like desktop app to generate private, AI-powered podcasts, in research preview across 20+ markets (Ivan Mehta/TechCrunch)

Model ReleasesDGX agent

Ivan Mehta / TechCrunch: Spotify Labs launches Studio, a NotebookLM-like desktop app to generate private, AI-powered podcasts, in research preview across 20+ markets — One of the common features for c

Spotify says it has 1M+ subscriptions to Audiobook+, which is on track for $100M in ARR, and unveils an ElevenLabs-powered audiobook self-publishing tool (Ivan Mehta/TechCrunch)

Model ReleasesDGX agent

Ivan Mehta / TechCrunch: Spotify says it has 1M+ subscriptions to Audiobook+, which is on track for $100M in ARR, and unveils an ElevenLabs-powered audiobook self-publishing tool — Alongside tools for

STAR-IOD: Scale-decoupled Topology Alignment with Pseudo-label Refinement for Remote Sensing Incremental Object Detection


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2605.20738v1 Announce Type: new Abstract: Remote sensing imagery typically arrives in the form of continuous data streams. Traditional detectors often forget previously learned categories when l

SymbolicLight V1: Spike-Gated Dual-Path Language Modeling with High Activation Sparsity and Sub-Billion-Scale Pre-Training Evidence

Model ReleasesDGX agent

arXiv:2605.21333v1 Announce Type: new Abstract: Natively trained spiking language models struggle to combine Transformer-like language quality, stable multi-domain pre-training, and high activation sp

TabPFN Extensions for Interpretable Geotechnical Modelling

Model ReleasesDGX agent

arXiv:2603.21033v2 Announce Type: replace-cross Abstract: Geotechnical site characterisation relies on sparse, heterogeneous borehole data, where uncertainty quantification and interpretability matter

TASTE: A Designer-Annotated Multi-Dimensional Preference Dataset for AI-Generated Graphic Design

Model ReleasesDGX agent

arXiv:2605.20731v1 Announce Type: new Abstract: Text-to-image models produce graphic design at production scale, but their supervision comes from photo-style preference data with a single overall verd

TempGlitch: Evaluating Vision-Language Models for Temporal Glitch Detection in Gameplay Videos

Model ReleasesDGX agent

arXiv:2605.21443v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly being explored for video game quality assurance, especially gameplay glitch detection. Most existing eval

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

Model ReleasesDGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

Text Analytics Evaluation Framework: A Case Study on LLMs and Social Media

Model ReleasesDGX agent

arXiv:2605.21338v1 Announce Type: new Abstract: LLMs have demonstrated exceptional proficiency in a wide range of NLP tasks. However, a notable gap remains in practical data analysis scenarios, partic

TextSculptor: Training and Benchmarking Scene Text Editing

Model ReleasesDGX agent

arXiv:2605.21090v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and diffusion-based generative models have substantially improved prompt-driven image editin

The Economics of Model Collapse: Equilibrium, Welfare, and Optimal Provenance Subsidies in Synthetic Data Markets

Model ReleasesDGX agent

arXiv:2605.20279v1 Announce Type: cross Abstract: Generative artificial intelligence is rapidly transforming the supply side of training data: an increasing share of new tokens, images, and structured

The top announcements for startups from Google I/O ‘26

Model ReleasesDGX agent

Many of the world’s fastest-growing AI startups are choosing to build their future — and the world’s — on Google Cloud because of our complete and open AI stack. We embed AI into every layer of our ar

The Visual Iconicity Challenge: Evaluating Vision-Language Models on Sign Language Form-Meaning Mapping

Model ReleasesDGX agent

arXiv:2510.08482v3 Announce Type: replace-cross Abstract: Iconicity, the resemblance between linguistic form and meaning, is pervasive in signed languages, offering a natural testbed for visual ground

The Yes-Man Syndrome: Benchmarking Abstention in Embodied Robotic Agents

Model ReleasesDGX agent

arXiv:2605.20544v1 Announce Type: cross Abstract: Vision-language models (VLMs) are used as high-level planners for embodied agents, translating natural language instructions and visual observations i

Theoretical guidelines for annealed Langevin dynamics in compositional simulation-based inference

Model ReleasesDGX agent

arXiv:2605.21253v1 Announce Type: cross Abstract: Compositional score-based approaches to simulation-based inference (SBI) approximate the posterior over a shared parameter given n independent observa

THEval. Evaluation Framework for Talking Head Video Generation

Model ReleasesDGX agent

arXiv:2511.04520v4 Announce Type: replace Abstract: Video generation has achieved remarkable progress, with generated videos increasingly resembling real ones. However, the rapid advance in generation

Time-Dependent PDE-Constrained Optimization via Weak-Form Latent Dynamics

Model ReleasesDGX agent

arXiv:2605.20639v1 Announce Type: cross Abstract: Optimization problems constrained by high-dimensional, time-dependent partial differential equations require repeated forward and sensitivity solves,

TimeSRL: Generalizable Time-Series Behavioral Modeling via Semantic RL-Tuned LLMs -- A Case Study in Mental Health

Model ReleasesDGX agent

arXiv:2605.21295v1 Announce Type: new Abstract: Longitudinal passive sensing enables continuous health prediction, yet models often fail under cross-dataset distribution shifts. Traditional ML overfit

Tippett-minimum Fusion of Representation-space Diffusion Models for Multi-Encoder Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2605.20502v1 Announce Type: cross Abstract: We address out-of-distribution (OOD) detection across the full spectrum of distribution shifts -- global domain changes, semantic divergence, texture

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million par…

Model ReleasesDGX agent

TLMs: Tiny LLMs and Agents on Edge Devices with @cormacb https://www.youtube.com/watch?v=-TiET_K-E_g Function Gemma ships at 270 million parameters and runs nearly 2,000 tokens per second prefill on a

To Select or not to Select, that is the Question: Distilling Robot Skill Prediction into a Small Ensemble

Model ReleasesDGX agent

arXiv:2605.21242v1 Announce Type: new Abstract: As robot fleets become more heterogeneous, including humanoids, rovers, quadrupeds, and drones, selecting the right robot for a task becomes a core syst

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefi…

Model ReleasesDGX agent

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefit one at a time inside a 1.7B byte-level pretraining pipelin

Token Maxxing, April 2026-May 2026, RIP 🪦

Model ReleasesDGX agent

Token Maxxing, April 2026-May 2026, RIP 🪦 🦔Microsoft canceled its internal Claude Code licenses this week after token-based billing made the cost untenable, even for a company with effectively infinit

Towards UAV Detection in the Real World: A New Multispectral Dataset UAVNet-MS and a New Method

Model ReleasesDGX agent

arXiv:2605.20963v1 Announce Type: new Abstract: The proliferation of unmanned aerial vehicles (UAVs) has created urgent demand for precise UAV monitoring. Existing RGB-based systems rely on spatial cu

Toxic Subword Pruning for Dialogue Response Generation on Large Language Models

Model ReleasesDGX agent

arXiv:2410.04155v2 Announce Type: replace Abstract: How to defend large language models (LLMs) from generating toxic content is an important research area. Yet, most research focused on various model

Training distribution determines the ceiling of drug-blind cancer sensitivity prediction

Model ReleasesDGX agent

arXiv:2605.20885v1 Announce Type: new Abstract: Precision oncology requires predicting which drugs will suppress a specific tumor from its molecular profile, but drug-blind sensitivity prediction has

Training Language Agents to Learn from Experience

Model ReleasesDGX agent

arXiv:2605.20477v1 Announce Type: cross Abstract: Language agents can adapt from experience in interactive environments, but current reflection-based methods can only self-correct within a single task

TRAM: Test-Time Risk Adaptation with Mixture of Agents

Model ReleasesDGX agent

arXiv:2408.08812v2 Announce Type: replace Abstract: Deployed reinforcement learning agents often face safety requirements that are specified only after training, such as new hazard maps, revised risk

Try Composer 2.5

Model ReleasesDGX agent

Try Composer 2.5 New CursorBench results just dropped. Two big takeaways. Composer 2.5 is way better than most people think. 63.2% score at $0.55 per task. Nearly matching Opus 4.7 Max and GPT 5.5 Ext

Tunable MAGMAX: Preference-Aware Model Merging for Continual Learning

Model ReleasesDGX agent

arXiv:2605.20803v1 Announce Type: new Abstract: Continual learning (CL) aims to train models sequentially on multiple tasks while mitigating catastrophic forgetting of previously learned knowledge. Re

Under Pressure: Emotional Framing Induces Measurable Behavioral Shifts and Structured Internal Geometry in Small Language Models

Model ReleasesDGX agent

arXiv:2605.20202v1 Announce Type: new Abstract: I study whether emotionally framed evaluation follow-ups change both the behavior and the calm-relative internal representations of small, locally deplo

Understanding and Improving Communication Performance in Multi-node LLM Inference

Model ReleasesDGX agent

arXiv:2511.09557v4 Announce Type: replace-cross Abstract: As large language models (LLMs) continue to grow in size, distributed inference has become increasingly important. Model-parallel strategies m

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs

Model ReleasesDGX agent

arXiv:2505.19075v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable general capabilities, but enhancing skills such as reasoning often demands substanti

VDFP: Video Deflickering with Flicker-banding Priors

Model ReleasesDGX agent

arXiv:2605.21079v1 Announce Type: new Abstract: Capturing digital screens with smartphones frequently induces severe banding due to hardware synchronization mismatches. Existing video restoration meth

Verifiable Provenance and Watermarking for Generative AI: An Evidentiary Framework for International Operational Law and Domestic Courts

Model ReleasesDGX agent

arXiv:2605.21002v1 Announce Type: cross Abstract: Generative artificial intelligence now synthesizes photorealistic imagery, audio, and video at a cost that defeats traditional forensic intuition. The

VersusQ: Pairwise Margin Reasoning for Generalizable Video Quality Assessment

Model ReleasesDGX agent

arXiv:2605.21130v1 Announce Type: new Abstract: Large Multimodal Models (LMMs) have shown promise for video quality assessment, but most methods still predict an absolute score for each video. Such po

Vision Transformers and Convolutional Neural Networks for Land Use Scene Classification

Model ReleasesDGX agent

arXiv:2605.21268v1 Announce Type: new Abstract: Land Use Scene Classification (LUSC) from remote sensing imagery plays a critical role in environmental monitoring, urban planning, and sustainable reso

VISTA: Technical Report for the Ego4D Short-Term Object Interaction Anticipation at EgoVis 2026

Model ReleasesDGX agent

arXiv:2605.20901v1 Announce Type: new Abstract: We propose VISTA, a V-JEPA Integrated StillFast Temporal Anticipator for the Ego4D Short-Term Object Interaction Anticipation (STA) Challenge at EgoVis

VISTAQA: Benchmarking Joint Visual Question Answering and Pixel-Level Evidence

Model ReleasesDGX agent

arXiv:2605.20676v1 Announce Type: new Abstract: Establishing a clear link between model predictions and the visual evidence that supports them is critical for transparency and reliability in multimoda

VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2605.20774v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong promise for general-purpose robotic manipulation, but their real-world evaluation remains limited

VSCD: Video-based Scene Change Detection in Unaligned Scenes

Model ReleasesDGX agent

arXiv:2605.20821v1 Announce Type: new Abstract: Detecting what has changed in an environment is essential for long-term autonomy, yet most change detection settings assume fixed viewpoints, mild misal

WaveGraphNet: Physics-Consistent Guided-Wave Damage Localization through Coupled Inverse-Forward Graph Learning

Model ReleasesDGX agent

arXiv:2605.20311v1 Announce Type: new Abstract: Guided-wave structural health monitoring enables damage localization in composite plates using sparse networks of bonded piezoelectric transducers. Howe

WCXB: A Multi-Type Web Content Extraction Benchmark

Model ReleasesDGX agent

arXiv:2605.21097v1 Announce Type: new Abstract: Web content extraction - isolating a page's main content from surrounding boilerplate - is a prerequisite for search indexing, retrieval-augmented gener

Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics

Model ReleasesDGX agent

arXiv:2605.20441v1 Announce Type: new Abstract: Transformers trained on modular arithmetic exhibit sharp transitions between memorization, generalization, and collapse. We show that weight decay acts

We’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental risks

Model ReleasesDGX agent

Google DeepMind has launched 'AI for the Planet,' a three-month accelerator program in Asia Pacific focused on leveraging advanced AI to combat environmental challenges like climate change, biodiversi

What Do Biomedical NER and Entity Linking Benchmarks Measure? A Corpus-Centric Diagnostic Framework

Model ReleasesDGX agent

arXiv:2605.20537v1 Announce Type: new Abstract: Biomedical named entity recognition (NER) and entity linking (EL) strongly depend on annotated corpora, but the utility of these resources for benchmark

What Twelve LLM Agent Benchmark Papers Disclose About Themselves: A Pilot Audit and an Open Scoring Schema

Model ReleasesDGX agent

arXiv:2605.21404v1 Announce Type: new Abstract: We read twelve well-known LLM agent benchmark papers and recorded, dimension by dimension, what each paper actually says about how its evaluation was ru

When Irregularity Helps: A Subclass Analysis of Inductive Bias in Neural Morphology

Model ReleasesDGX agent

arXiv:2605.20558v1 Announce Type: new Abstract: Neural morphological generation systems often achieve high aggregate accuracy on benchmark datasets, yet such performance can conceal systematic errors

When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size Sufficiency

Model ReleasesDGX agent

arXiv:2603.09024v2 Announce Type: replace Abstract: Sudden concept drift makes previously trained predictors unreliable, yet deciding when to retrain and what post-drift data size is sufficient is rar

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata

Model ReleasesDGX agent

arXiv:2605.21479v1 Announce Type: new Abstract: Visual Question Answering (VQA) benchmarks have largely emphasized perception-based tasks that can be solved from visual content alone. In contrast, man

WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents

Model ReleasesDGX agent

arXiv:2605.20306v1 Announce Type: new Abstract: We introduce WildRoadBench, a wild aerial road-damage grounding benchmark that couples direct visual grounding by vision-language models with autonomous

Winfree Oscillatory Neural Network

Model ReleasesDGX agent

arXiv:2605.20922v1 Announce Type: cross Abstract: Oscillations and synchronization are widely believed to play a fundamental role in representation and computation. However, existing machine learning

Wordle 1,796 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟩⬛🟨 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post shows a Wordle game result where the player solved puzzle #1,796 in 4 attempts, with the final answer being a five-letter word with the pattern shown in green squares. The emoji grid display

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories

Model ReleasesDGX agent

arXiv:2605.21468v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving reasoning in large language models (LLMs), yet the

ZEBRA: Zero-shot Budgeted Resource Allocation for LLM Orchestration

Model ReleasesDGX agent

arXiv:2605.20485v1 Announce Type: new Abstract: As autonomous agents increasingly execute end-to-end tasks under fixed monetary budgets, the pressing open question shifts from whether the budget is re

20 May 2026

100 things we announced at I/O 2026

Model ReleasesDGX agent

Google I/O 2026 unveiled new models, agents and tools to help users build, search, create, discover, shop and get more done. Key announcements included Gemini Omni, Google Antigravity, and Universal C

A Bitter Lesson for Data Filtering

Model ReleasesDGX agent

arXiv:2605.19407v1 Announce Type: cross Abstract: We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an app

A Case for Agentic Tuning: From Documentation to Action in PostgreSQL

Model ReleasesDGX agent

arXiv:2605.19988v1 Announce Type: cross Abstract: Documentation has long guided computer system tuning by distilling expert knowledge into per-parameter recommendations. Yet such guides capture only w

A Family of Divergence Measures for Evaluating the Reconstruction Quality of Explainable Ensemble Trees

Model ReleasesDGX agent

arXiv:2605.19618v1 Announce Type: new Abstract: Validating interpretable surrogate models for ensemble learners requires measuring agreement between the ensemble's internal representation and its surr

A Hybrid Modeling Framework for Crop Prediction Tasks via Dynamic Parameter Calibration and Multi-Task Learning

Model ReleasesDGX agent

arXiv:2603.15411v2 Announce Type: replace Abstract: Accurate prediction of crop states (e.g., phenology stages and cold hardiness) is essential for timely farm management decisions such as irrigation,

← Previous
1…225226227228229…377
Next →