AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,585 results
20 May 2026

A Nonlinear Complexity Index for Wearable PPG Cardiovascular Stability: Multiscale Validation, Systematic Evaluation Correction, and Bayesian Parameter Optimization

Model ReleasesDGX agent

arXiv:2605.18802v1 Announce Type: cross Abstract: Cardiovascular stability estimation from wearable photoplethysmography (PPG) requires a principled nonlinear framework, yet major gaps persist in heur

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

Model ReleasesDGX agent

arXiv:2605.18780v1 Announce Type: cross Abstract: Reasoning-based Large Language Models (LLMs) like PO4ISR have set new benchmarks in session-based recommendation. However, the reproducibility of thei

A Systematic Failure Analysis of Vision Foundation Models for Open Set Iris Presentation Attack Detection


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2605.19020v1 Announce Type: new Abstract: Vision foundation models have demonstrated strong transferability across diverse visual recognition tasks and are increasingly considered for biometric

A Two-Parameter Weibull Framework for Diagnosing Transformer Weight Distributions

Model ReleasesDGX agent

arXiv:2605.18898v1 Announce Type: new Abstract: We apply the Weibull distribution -- a two-parameter family from extreme-value theory -- as a diagnostic framework for element-wise weight magnitude dis

Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment

Model ReleasesDGX agent

arXiv:2506.14148v2 Announce Type: replace-cross Abstract: This paper presents a novel non-invasive object classification approach using acoustic scattering, demonstrated through a case study on hair a

Active Learning of Fractional-Order Viscoelastic Model Parameters for Realistic Haptic Rendering

Model ReleasesDGX agent

arXiv:2512.00667v2 Announce Type: replace-cross Abstract: Effective medical simulators necessitate realistic haptic rendering of biological tissues that exhibit viscoelastic material properties, such

Adapted Center and Scale Prediction: More Stable and More Accurate

Model ReleasesDGX agent

arXiv:2002.09053v3 Announce Type: replace Abstract: Pedestrian detection benefits from deep learning technology and gains rapid development in recent years. Most of detectors follow general object det

Adaptive Power Iteration Method for Differentially Private PCA

Model ReleasesDGX agent

arXiv:2602.11454v3 Announce Type: replace-cross Abstract: We study left(epsilon,eltaright)-differentially private algorithms for the problem of approximately computing the top singular vector of a mat

Add a Specialized Deep Research Skill to Agent Harnesses

Model ReleasesDGX agent

This article covers the NVIDIA AI-Q Blueprint for building specialized deep research agents that empower AI systems to gather context, synthesize information, and support complex decision-making acros

Addressing prior dependence in hierarchical Bayesian modeling for PTA data analysis II: Noise and SGWB inference through parameter decorrelation

Model ReleasesDGX agent

arXiv:2511.01959v2 Announce Type: replace-cross Abstract: Pulsar Timing Arrays (PTA) provide a powerful framework to measure low-frequency gravitational waves, but accuracy and robustness of the resul

Adversarial Stress Testing of SPARK Humanoid Safety Filters

Model ReleasesDGX agent

arXiv:2605.19009v1 Announce Type: new Abstract: Humanoid robots are difficult to deploy safely because they have high-dimensional bodies, many collision constraints, and must operate near people and o

Aero-World: Action-Conditioned Aerial Video Generation from Inertial Controls

Model ReleasesDGX agent

arXiv:2605.19728v1 Announce Type: new Abstract: Foundation video models produce visually impressive results, but their use in embodied AI remains limited because they are primarily trained on natural

Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents

Model ReleasesDGX agent

arXiv:2605.19149v1 Announce Type: new Abstract: Agents operating with computer and Web use inevitably encounter errors: inaccessible webpages, missing files, local and remote misconfigurations, etc. T

AgentNLQ: A General-Purpose Agent for Natural Language to SQL

Model ReleasesDGX agent

arXiv:2605.19010v1 Announce Type: new Abstract: Natural language to SQL (NL2SQL) conversion is an important problem for researchers and enterprises due to the ubiquitous importance of relational datab

[AINews] Google I/O 2026: Gemini 3.5 Flash, Omni (NanoBanana for Video), Spark (background agents), and Antigravity 2.0

Model ReleasesDGX agent

Google I/O 2026 featured several new AI model releases including Gemini 3.5 Flash, an Omni model codenamed NanoBanana for video processing, Spark for background agent tasks, and Antigravity 2.0. These

An Exterior Method for Nonnegative Matrix Factorization

Model ReleasesDGX agent

arXiv:2605.19325v1 Announce Type: new Abstract: Nonnegative matrix factorization (NMF) seeks a low-rank approximation X approx UV^T with nonnegative factors and is commonly solved using interior metho

An interview with Match Group CEO Spencer Rascoff about plans for Tinder, including a redesign, AI features, live events, and group dating to win over Gen Z (Samantha Kelly/Bloomberg)

Model ReleasesDGX agent

Samantha Kelly / Bloomberg: An interview with Match Group CEO Spencer Rascoff about plans for Tinder, including a redesign, AI features, live events, and group dating to win over Gen Z — The dating ap

An LLM-Based System for Argument Mining

Model ReleasesDGX agent

arXiv:2605.13793v2 Announce Type: replace Abstract: Arguments are a fundamental aspect of human reasoning, in which claims are supported, challenged, and weighed against one another. We present an end

An OpenAI model has disproved a central conjecture in discrete geometry

Model ReleasesDGX agent

An OpenAI AI model successfully disproved a longstanding conjecture in discrete geometry, a mathematical field studying geometric properties of discrete objects. This achievement demonstrates the pote

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term t…

Model ReleasesDGX agent

Anyone understand what Google mean by 'Gemini Spark runs on Gemini 3.5 and uses the Antigravity harness' - is 'Antigravity' a generic term they're using for their agent harnesses now or is their Claw-

Are Tools Always Beneficial? Learning to Invoke Tools Adaptively for Dual-Mode Multimodal LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.19852v1 Announce Type: new Abstract: Tool-augmented reasoning has emerged as a promising direction for enhancing the reasoning capabilities of multimodal large language models (MLLMs). Howe

Artifact-Bench: Evaluating MLLMs on Detecting and Assessing the Artifacts of AI-Generated Videos

Model ReleasesDGX agent

arXiv:2605.18984v1 Announce Type: new Abstract: Recent video generative models have greatly improved the realism of AI-generated videos, yet their outputs still exhibit artifacts such as temporal inco

Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

Model ReleasesDGX agent

arXiv:2605.18891v1 Announce Type: cross Abstract: Evaluations of unlearning on reasoning models sometimes show a bypass pattern. The answer side looks unlearned, but the model's own thinking trace kee

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

Model ReleasesDGX agent

arXiv:2605.20025v1 Announce Type: new Abstract: Automating scientific discovery requires more than generating papers from ideas. Real research is iterative: hypotheses are challenged from multiple per

Backdooring Masked Diffusion Language Models

Model ReleasesDGX agent

arXiv:2605.19262v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) are emerging as a compelling new paradigm for text generation, but their training-time security remains largely

Base Models Look Human To AI Detectors

Model ReleasesDGX agent

arXiv:2605.19516v1 Announce Type: cross Abstract: As AI-generated text enters the real-world at scale, institutions increasingly use commercial AI-text detectors, especially in education and academic-

Benchmark and optimize LLMs on-device with AI Edge Portal

Model ReleasesDGX agent

LLMs have become more powerful at smaller sizes, but deploying them to edge devices like smartphones remains a massive challenge. Today, developers have to optimize across a sprawling combination of a

Benchmarking and Evolving Reason-Reflect-Rectify for Reflective Visual Generation

Model ReleasesDGX agent

arXiv:2605.19639v1 Announce Type: new Abstract: Text-to-Image (T2I) models and Unified Multimodal Models (UMMs) have achieved remarkable progress in visual generation. However, their reliance on a sin

Benchmarking Commercial ASR Systems on Code-Switching Speech: Arabic, Persian, and German

Model ReleasesDGX agent

arXiv:2605.19069v1 Announce Type: cross Abstract: Code-switching -- the natural alternation between two languages within a single utterance -- represents one of the most challenging and under-studied

Beyond Binary Success: A Diagnostic Meta-Evaluation Framework for Fine-Grained Manipulation

Model ReleasesDGX agent

arXiv:2605.19986v1 Announce Type: cross Abstract: Fine-grained manipulation marks a regime where global scene context no longer suffices, and success hinges on the tight coupling of local attribute gr

Beyond Imitation: Learning Safe End-to-End Autonomous Driving from Hard Negatives

Model ReleasesDGX agent

arXiv:2605.19771v1 Announce Type: cross Abstract: Existing imitation learning methods for end-to-end autonomous driving predominantly learn from successful demonstrations by minimizing geometric devia

Beyond Prediction Accuracy: Target-Space Recovery Profiles for Evaluating Model-Brain Alignment

Model ReleasesDGX agent

arXiv:2605.20127v1 Announce Type: cross Abstract: Artificial vision models are often evaluated against the human visual cortex by measuring how accurately their internal representations predict brain

Beyond Waypoints: Dual-Heatmap Grounding for Cross-Embodiment Semantic Navigation

Model ReleasesDGX agent

arXiv:2605.19420v1 Announce Type: new Abstract: Grounding open-ended semantic instructions into physically executable local goals is a fundamental challenge in human-robot interaction. While existing

BLINKG: A Benchmark for LLM-Integrated Knowledge Graph Generation

Model ReleasesDGX agent

arXiv:2605.19518v1 Announce Type: new Abstract: Generating Knowledge Graphs (KGs) remains one of the most time-consuming and labor-intensive tasks for knowledge engineers, as they need to identify sem

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and …

Model ReleasesDGX agent

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and encounter. They’re not really random benchmark tasks, they r

BuildArena: A Physics-Aligned Interactive Benchmark of LLMs for Engineering Construction

Model ReleasesDGX agent

arXiv:2510.16559v5 Announce Type: replace Abstract: Engineering construction automation aims to transform natural language specifications into physically viable structures, requiring complex integrate

Can Large Language Models Reliably Correct Errors in Low-Resource ASR? A Contamination-Aware Case Study on West Frisian

Model ReleasesDGX agent

arXiv:2605.19711v1 Announce Type: new Abstract: Automatic speech recognition (ASR) has improved substantially in recent years, yet performance remains limited for low-resource languages. Large languag

Can LLMs Emulate Human Belief Dynamics?

Model ReleasesDGX agent

arXiv:2605.18781v1 Announce Type: cross Abstract: Can LLMs simulate how humans form and change beliefs in social networks? We put this to the test by replicating an established study on belief dynamic

CaptchaMind: Training CAPTCHA Solvers via Reinforcement Learning with Explicit Reasoning Supervision

Model ReleasesDGX agent

arXiv:2605.19538v1 Announce Type: cross Abstract: CAPTCHAs are widely deployed as human verification mechanisms and frequently block intelligent agents from completing end-to-end automation in real-wo

Causal Evidence for Attention Head Imbalance in Modality Conflict Hallucination

Model ReleasesDGX agent

arXiv:2605.19250v1 Announce Type: new Abstract: Modality-conflict hallucination occurs when multimodal large language models (MLLMs) prioritize erroneous textual premises over contradictory visual evi

Chunking German Legal Code

Model ReleasesDGX agent

arXiv:2605.19806v1 Announce Type: cross Abstract: This paper investigates chunking strategies for retrieval-augmented generation on German statutory law, using the German Civil Code as a structured be

ClinSeekAgent: Automating Multimodal Evidence Seeking for Agentic Clinical Reasoning

Model ReleasesDGX agent

arXiv:2605.20176v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems have shown promise for clinical decision support, but existing works largely assume that evidence has a

ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation

Model ReleasesDGX agent

arXiv:2605.18769v1 Announce Type: cross Abstract: Personalized Retrieval-Augmented Generation (RAG) relies on accurately selecting user-relevant documents. In practice, existing RAG approaches often s

CogScale: Scalable Benchmark for Sequence Processing

Model ReleasesDGX agent

arXiv:2605.19758v1 Announce Type: new Abstract: The ability to maintain and manipulate information over time is a fundamental aspect of living beings and Artificial Intelligence. While modern models h

COMPASS: Confined-space Manipulation Planning with Active Sensing Strategy

Model ReleasesDGX agent

arXiv:2509.14787v2 Announce Type: replace Abstract: Manipulation in confined and cluttered environments remains a significant challenge due to partial observability and complex configuration spaces. E

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

Model ReleasesDGX agent

arXiv:2605.18808v1 Announce Type: cross Abstract: We characterize a compositional architecture of literary primitives in two instruction-tuned large language models (Llama 3.1 8B-Instruct and Gemma 2

Conflict-Resilient Multi-Agent Reasoning via Signed Graph Modeling

Model ReleasesDGX agent

arXiv:2605.19418v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) have demonstrated strong reasoning and decision-making capabilities that consistently surpass those of single LLM ag

Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations

Model ReleasesDGX agent

arXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga

CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering

Model ReleasesDGX agent

arXiv:2605.19075v1 Announce Type: cross Abstract: Grounded multi-video question answering over real-world news events requires systems to surface query-relevant evidence across heterogeneous video arc

Cross-View Attention Fusion Net: A Prior-Guided Dual-View Representation Learning for Cardiac Output Estimation from Short-Term PPG Signals

Model ReleasesDGX agent

arXiv:2605.19666v1 Announce Type: cross Abstract: Accurate cardiac output (CO) estimation from photoplethysmography (PPG) is promising for unobtrusive hemodynamic monitoring, but remains difficult sin

Cross-View Splatter: Feed-Forward View Synthesis with Georeferenced Images

Model ReleasesDGX agent

arXiv:2605.19656v1 Announce Type: new Abstract: We present Cross-View Splatter, a feed-forward method that predicts pixel-aligned Gaussian splats for outdoor scenes captured at ground level AND by sat

CutVerse: A Compositional GUI Agents Benchmark for Media Post-Production Editing

Model ReleasesDGX agent

arXiv:2605.19484v1 Announce Type: cross Abstract: While GUI agents have made significant progress in web navigation and basic operating system tasks, their capabilities in professional creative workfl

D^3-Subsidy: Online and Sequential Driver Subsidy Decision-Making for Large-Scale Ride-Hailing Market

Model ReleasesDGX agent

arXiv:2605.20036v1 Announce Type: new Abstract: Ride-hailing platforms like DiDi Chuxing operate in highly dynamic environments where balancing driver supply and passenger demand is critical. Although

deadtrees.earth-aerial: A Multi-Resolution Aerial Image Dataset for Tree Cover and Mortality Detection

Model ReleasesDGX agent

arXiv:2605.19605v1 Announce Type: new Abstract: Forests worldwide are increasingly threatened by climate change and disturbances such as fire, pests, and pathogens, creating an urgent need for scalabl

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows

Model ReleasesDGX agent

arXiv:2605.19099v1 Announce Type: new Abstract: We introduce DecisionBench, a benchmark substrate for emergent delegation in long-horizon agentic workflows. The substrate fixes a task suite (GAIA, tau

Depth2Pose: A Pose-Based Benchmark for Monocular Depth Estimation without Ground-Truth Depth

Model ReleasesDGX agent

arXiv:2605.19797v1 Announce Type: new Abstract: Monocular depth estimation has improved significantly in recent years, driven by increasingly powerful models and large-scale training data. Predicted d

Detecting Fluent Optimization-Based Adversarial Prompts via Sequential Entropy Changes

Model ReleasesDGX agent

arXiv:2605.19966v1 Announce Type: cross Abstract: Optimization-based adversarial suffixes can jailbreak aligned large language models (LLMs) while remaining fluent, weakening static and windowed perpl

Did we ever learn what model won gold at the IMO from OpenAI? It was a year ago and it was called an unreleased internal general purpose mod…

Model ReleasesDGX agent

Did we ever learn what model won gold at the IMO from OpenAI? It was a year ago and it was called an unreleased internal general purpose model back then. Has GPT-5.5 Pro Extended caught up with whatev

Differential-Integral Neural Operator for Long-Term Turbulence Forecasting

Model ReleasesDGX agent

arXiv:2509.21196v3 Announce Type: replace-cross Abstract: Accurately forecasting the long-term evolution of turbulence represents a grand challenge in scientific computing and is crucial for applicati

Disentangling generalization and memorization in large language models using chess

Model ReleasesDGX agent

arXiv:2601.16823v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) exhibit remarkable capabilities, yet it remains unclear to what extent these reflect sophisticated recall or genu

← Previous
1…226227228229230…377
Next →