AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

VISTA: Technical Report for the Ego4D Short-Term Object Interaction Anticipation at EgoVis 2026

DGX agent

arXiv:2605.20901v1 Announce Type: new Abstract: We propose VISTA, a V-JEPA Integrated StillFast Temporal Anticipator for the Ego4D Short-Term Object Interaction Anticipation (STA) Challenge at EgoVis

model-releasesarxiv-cs-cv
21 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

VISTAQA: Benchmarking Joint Visual Question Answering and Pixel-Level Evidence

DGX agent

arXiv:2605.20676v1 Announce Type: new Abstract: Establishing a clear link between model predictions and the visual evidence that supports them is critical for transparency and reliability in multimoda

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

VLA-REPLICA: A Low-Cost, Reproducible Benchmark for Real-World Evaluation of Vision-Language-Action Models

DGX agent

arXiv:2605.20774v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong promise for general-purpose robotic manipulation, but their real-world evaluation remains limited

model-releasesarxiv-cs-ro
21 May 2026
Model Releases

VSCD: Video-based Scene Change Detection in Unaligned Scenes

DGX agent

arXiv:2605.20821v1 Announce Type: new Abstract: Detecting what has changed in an environment is essential for long-term autonomy, yet most change detection settings assume fixed viewpoints, mild misal

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

WaveGraphNet: Physics-Consistent Guided-Wave Damage Localization through Coupled Inverse-Forward Graph Learning

DGX agent

arXiv:2605.20311v1 Announce Type: new Abstract: Guided-wave structural health monitoring enables damage localization in composite plates using sparse networks of bonded piezoelectric transducers. Howe

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

WCXB: A Multi-Type Web Content Extraction Benchmark

DGX agent

arXiv:2605.21097v1 Announce Type: new Abstract: Web content extraction - isolating a page's main content from surrounding boilerplate - is a prerequisite for search indexing, retrieval-augmented gener

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics

DGX agent

arXiv:2605.20441v1 Announce Type: new Abstract: Transformers trained on modular arithmetic exhibit sharp transitions between memorization, generalization, and collapse. We show that weight decay acts

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

We’re launching the Google DeepMind Accelerator program in Asia Pacific to tackle environmental risks

DGX agent

Google DeepMind has launched 'AI for the Planet,' a three-month accelerator program in Asia Pacific focused on leveraging advanced AI to combat environmental challenges like climate change, biodiversi

model-releasesgoogle-deepmind
21 May 2026
Model Releases

What Do Biomedical NER and Entity Linking Benchmarks Measure? A Corpus-Centric Diagnostic Framework

DGX agent

arXiv:2605.20537v1 Announce Type: new Abstract: Biomedical named entity recognition (NER) and entity linking (EL) strongly depend on annotated corpora, but the utility of these resources for benchmark

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

What Twelve LLM Agent Benchmark Papers Disclose About Themselves: A Pilot Audit and an Open Scoring Schema

DGX agent

arXiv:2605.21404v1 Announce Type: new Abstract: We read twelve well-known LLM agent benchmark papers and recorded, dimension by dimension, what each paper actually says about how its evaluation was ru

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

When Irregularity Helps: A Subclass Analysis of Inductive Bias in Neural Morphology

DGX agent

arXiv:2605.20558v1 Announce Type: new Abstract: Neural morphological generation systems often achieve high aggregate accuracy on benchmark datasets, yet such performance can conceal systematic errors

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

When to Retrain after Drift: A Data-Only Test of Post-Drift Data Size Sufficiency

DGX agent

arXiv:2603.09024v2 Announce Type: replace Abstract: Sudden concept drift makes previously trained predictors unreliable, yet deciding when to retrain and what post-drift data size is sufficient is rar

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata

DGX agent

arXiv:2605.21479v1 Announce Type: new Abstract: Visual Question Answering (VQA) benchmarks have largely emphasized perception-based tasks that can be solved from visual content alone. In contrast, man

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents

DGX agent

arXiv:2605.20306v1 Announce Type: new Abstract: We introduce WildRoadBench, a wild aerial road-damage grounding benchmark that couples direct visual grounding by vision-language models with autonomous

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Winfree Oscillatory Neural Network

DGX agent

arXiv:2605.20922v1 Announce Type: cross Abstract: Oscillations and synchronization are widely believed to play a fundamental role in representation and computation. However, existing machine learning

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Wordle 1,796 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟩⬛🟨 🟩🟩🟩🟩🟩

DGX agent

This post shows a Wordle game result where the player solved puzzle #1,796 in 4 attempts, with the final answer being a five-letter word with the pattern shown in green squares. The emoji grid display

model-releasesanthropic--x
21 May 2026
Model Releases

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories

DGX agent

arXiv:2605.21468v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a dominant paradigm for improving reasoning in large language models (LLMs), yet the

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

ZEBRA: Zero-shot Budgeted Resource Allocation for LLM Orchestration

DGX agent

arXiv:2605.20485v1 Announce Type: new Abstract: As autonomous agents increasingly execute end-to-end tasks under fixed monetary budgets, the pressing open question shifts from whether the budget is re

model-releasesarxiv-cs-lg
21 May 2026
Model Releases

100 things we announced at I/O 2026

DGX agent

Google I/O 2026 unveiled new models, agents and tools to help users build, search, create, discover, shop and get more done. Key announcements included Gemini Omni, Google Antigravity, and Universal C

model-releasesgoogle-ai
20 May 2026
Model Releases

A Bitter Lesson for Data Filtering

DGX agent

arXiv:2605.19407v1 Announce Type: cross Abstract: We investigate data filtering for large model pretraining via new scaling studies that target the high compute, data-scarce regime. In spite of an app

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Case for Agentic Tuning: From Documentation to Action in PostgreSQL

DGX agent

arXiv:2605.19988v1 Announce Type: cross Abstract: Documentation has long guided computer system tuning by distilling expert knowledge into per-parameter recommendations. Yet such guides capture only w

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Family of Divergence Measures for Evaluating the Reconstruction Quality of Explainable Ensemble Trees

DGX agent

arXiv:2605.19618v1 Announce Type: new Abstract: Validating interpretable surrogate models for ensemble learners requires measuring agreement between the ensemble's internal representation and its surr

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

A Hybrid Modeling Framework for Crop Prediction Tasks via Dynamic Parameter Calibration and Multi-Task Learning

DGX agent

arXiv:2603.15411v2 Announce Type: replace Abstract: Accurate prediction of crop states (e.g., phenology stages and cold hardiness) is essential for timely farm management decisions such as irrigation,

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Nonlinear Complexity Index for Wearable PPG Cardiovascular Stability: Multiscale Validation, Systematic Evaluation Correction, and Bayesian Parameter Optimization

DGX agent

arXiv:2605.18802v1 Announce Type: cross Abstract: Cardiovascular stability estimation from wearable photoplethysmography (PPG) requires a principled nonlinear framework, yet major gaps persist in heur

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

DGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Systematic Failure Analysis of Vision Foundation Models for Open Set Iris Presentation Attack Detection

DGX agent

arXiv:2605.19020v1 Announce Type: new Abstract: Vision foundation models have demonstrated strong transferability across diverse visual recognition tasks and are increasingly considered for biometric

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

A Two-Parameter Weibull Framework for Diagnosing Transformer Weight Distributions

DGX agent

arXiv:2605.18898v1 Announce Type: new Abstract: We apply the Weibull distribution -- a two-parameter family from extreme-value theory -- as a diagnostic framework for element-wise weight magnitude dis

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment

DGX agent

arXiv:2506.14148v2 Announce Type: replace-cross Abstract: This paper presents a novel non-invasive object classification approach using acoustic scattering, demonstrated through a case study on hair a

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Active Learning of Fractional-Order Viscoelastic Model Parameters for Realistic Haptic Rendering

DGX agent

arXiv:2512.00667v2 Announce Type: replace-cross Abstract: Effective medical simulators necessitate realistic haptic rendering of biological tissues that exhibit viscoelastic material properties, such

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Adapted Center and Scale Prediction: More Stable and More Accurate

DGX agent

arXiv:2002.09053v3 Announce Type: replace Abstract: Pedestrian detection benefits from deep learning technology and gains rapid development in recent years. Most of detectors follow general object det

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Adaptive Power Iteration Method for Differentially Private PCA

DGX agent

arXiv:2602.11454v3 Announce Type: replace-cross Abstract: We study left(epsilon,eltaright)-differentially private algorithms for the problem of approximately computing the top singular vector of a mat

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Add a Specialized Deep Research Skill to Agent Harnesses

DGX agent

This article covers the NVIDIA AI-Q Blueprint for building specialized deep research agents that empower AI systems to gather context, synthesize information, and support complex decision-making acros

model-releasesnvidia-developer
20 May 2026
Model Releases

Addressing prior dependence in hierarchical Bayesian modeling for PTA data analysis II: Noise and SGWB inference through parameter decorrelation

DGX agent

arXiv:2511.01959v2 Announce Type: replace-cross Abstract: Pulsar Timing Arrays (PTA) provide a powerful framework to measure low-frequency gravitational waves, but accuracy and robustness of the resul

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Adversarial Stress Testing of SPARK Humanoid Safety Filters

DGX agent

arXiv:2605.19009v1 Announce Type: new Abstract: Humanoid robots are difficult to deploy safely because they have high-dimensional bodies, many collision constraints, and must operate near people and o

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls

DGX agent

arXiv:2605.19728v1 Announce Type: new Abstract: Foundation video models produce visually impressive results, but their use in embodied AI remains limited because they are primarily trained on natural

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents

DGX agent

arXiv:2605.19149v1 Announce Type: new Abstract: Agents operating with computer and Web use inevitably encounter errors: inaccessible webpages, missing files, local and remote misconfigurations, etc. T

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

AgentNLQ: A General-Purpose Agent for Natural Language to SQL

DGX agent

arXiv:2605.19010v1 Announce Type: new Abstract: Natural language to SQL (NL2SQL) conversion is an important problem for researchers and enterprises due to the ubiquitous importance of relational datab

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

[AINews] Google I/O 2026: Gemini 3.5 Flash, Omni (NanoBanana for Video), Spark (background agents), and Antigravity 2.0

DGX agent

Google I/O 2026 featured several new AI model releases including Gemini 3.5 Flash, an Omni model codenamed NanoBanana for video processing, Spark for background agent tasks, and Antigravity 2.0. These

model-releaseslatent-space
20 May 2026
Model Releases

An Exterior Method for Nonnegative Matrix Factorization

DGX agent

arXiv:2605.19325v1 Announce Type: new Abstract: Nonnegative matrix factorization (NMF) seeks a low-rank approximation X approx UV^T with nonnegative factors and is commonly solved using interior metho

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

An interview with Match Group CEO Spencer Rascoff about plans for Tinder, including a redesign, AI features, live events, and group dating to win over Gen Z (Samantha Kelly/Bloomberg)

DGX agent

Samantha Kelly / Bloomberg: An interview with Match Group CEO Spencer Rascoff about plans for Tinder, including a redesign, AI features, live events, and group dating to win over Gen Z — The dating ap

model-releasestechmeme
20 May 2026
Model Releases

An LLM-Based System for Argument Mining

DGX agent

arXiv:2605.13793v2 Announce Type: replace Abstract: Arguments are a fundamental aspect of human reasoning, in which claims are supported, challenged, and weighed against one another. We present an end

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

An OpenAI model has disproved a central conjecture in discrete geometry

DGX agent

An OpenAI AI model successfully disproved a longstanding conjecture in discrete geometry, a mathematical field studying geometric properties of discrete objects. This achievement demonstrates the pote

model-releasesopenai
20 May 2026
Model Releases

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term t…

DGX agent

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term they're using for their agent harnesses now or is their Claw-

model-releasessimon-willison--x
20 May 2026
Model Releases

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning

DGX agent

arXiv:2605.19852v1 Announce Type: new Abstract: Tool-augmented reasoning has emerged as a promising direction for enhancing the reasoning capabilities of multimodal large language models (MLLMs). Howe

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos

DGX agent

arXiv:2605.18984v1 Announce Type: new Abstract: Recent video generative models have greatly improved the realism of AI-generated videos, yet their outputs still exhibit artifacts such as temporal inco

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

DGX agent

arXiv:2605.18891v1 Announce Type: cross Abstract: Evaluations of unlearning on reasoning models sometimes show a bypass pattern. The answer side looks unlearned, but the model's own thinking trace kee

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

DGX agent

arXiv:2605.20025v1 Announce Type: new Abstract: Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple per

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Backdooring Masked Diffusion Language Models

DGX agent

arXiv:2605.19262v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are emerging as a compelling new paradigm for text generation, but their training-time security remains largely

model-releasesarxiv-cs-lg
20 May 2026
← Previous
1…284285286287288…472
Next →