AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Safety

TARIC: Memory-Augmented Traversability-Aware Outdoor VLN under Interrupted Semantic Cues

DGX agent

arXiv:2605.31121v1 Announce Type: cross Abstract: Outdoor vision-language navigation (VLN) in long-range, open-world environments is frequently disrupted by semantic-cue interruptions, where informati

safetyarxiv-cs-ai
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability

DGX agent

arXiv:2605.30628v1 Announce Type: cross Abstract: Universal LLM reliability is not a finite-library problem: across all possible tasks, tools, schemas, knowledge sources, and evaluator expectations, n

local-aiarxiv-cs-ai
1 Jun 2026
Safety

Zero Collapse: A Failure Mode of Policy Gradient Methods in Discontinuous Reward Environments

DGX agent

arXiv:2605.30896v1 Announce Type: new Abstract: Bidding in repeated auctions is a central challenge for reinforcement learning (RL), combining continuous control with the strategic complexities of dig

safetyarxiv-cs-lg
1 Jun 2026
Safety

Automating Low-Risk Code Review at Meta: RADAR, Risk Calibration, and Review Efficiency

DGX agent

arXiv:2605.30208v1 Announce Type: cross Abstract: AI-assisted coding tools have altered software production. At Meta, significant lines of code per human-landed diff grew by 105.9% year over year and

safetyarxiv-cs-ai
29 May 2026
Model Releases

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

DGX agent

arXiv:2509.23571v3 Announce Type: replace-cross Abstract: As cyber threats continue to grow in scale and sophistication, blue team defenders increasingly require advanced tools to proactively detect a

model-releasesarxiv-cs-ai
29 May 2026
Safety

Causal-JEPA: Learning World Models through Object-Level Latent Masking

DGX agent

arXiv:2602.11389v2 Announce Type: replace Abstract: World models require robust relational understanding to support prediction, reasoning, and control. While object-centric representations provide a u

safetyarxiv-cs-ai
29 May 2026
Safety

Certified Policy Optimisation for Nested Causal Bandits via PAC-Bayes Risk

DGX agent

arXiv:2605.29788v1 Announce Type: new Abstract: Critical sequential decisions are rarely single-timescale: a strategic decision causally shapes the context in which every subsequent tactical choice is

safetyarxiv-cs-ai
29 May 2026
Research

Conf-Gen: Conformal Uncertainty Quantification for Generative Models

DGX agent

arXiv:2605.28920v1 Announce Type: cross Abstract: Conformal prediction (CP) and its extension, conformal risk control (CRC), are established frameworks for quantifying uncertainty in supervised machin

researcharxiv-cs-ai
29 May 2026
Model Releases

Cookie-Bench: Continuous On-screen Key Interaction Evaluation for Web Generation

DGX agent

arXiv:2605.30000v1 Announce Type: new Abstract: Front-end web code has become a core product surface for every frontier LLM release, yet evaluating these interactive applications at development speed

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

DiffSpot: Can VLMs Spot Fine-Grained Visual Differences in Web Interfaces?

DGX agent

arXiv:2605.29615v1 Announce Type: cross Abstract: Vision-language models (VLMs) have made strong progress on high-level image-text alignment, yet their ability to perceive subtle visual differences re

model-releasesarxiv-cs-cl
29 May 2026
Safety

Emergent Semantic Representations in World Models through Physical Interaction without Linguistic Supervision

DGX agent

arXiv:2605.28865v1 Announce Type: cross Abstract: What does a world model learn from physical exploration, without any linguistic supervision? We argue the answer is organized by a single principle: t

safetyarxiv-cs-ai
29 May 2026
Model Releases

GrowLoop: Self-Evolving Conversation Evaluation Seeded by Human

DGX agent

arXiv:2605.28882v1 Announce Type: cross Abstract: With the rapid advancement of large language models, evaluating human-likeness in open-ended conversation has become increasingly important. However,

model-releasesarxiv-cs-ai
29 May 2026
Safety

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

DGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

safetyarxiv-cs-ai
29 May 2026
Tutorials

MEMENTO: Leveraging Web as a Learning Signal for Low-Data Domains

DGX agent

arXiv:2605.29795v1 Announce Type: new Abstract: Real-world tasks often lack large labeled datasets, motivating extensive work on learning in low-data regimes. However, existing approaches such as few-

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

Minimal Prompt Perturbations Lead to Code Vulnerabilities: Prompt Fragility and Hidden-State Signals in Coding LLMs

DGX agent

arXiv:2605.29737v1 Announce Type: cross Abstract: LLM-based coding assistants are seeing rapid adoption, offering substantial gains in developer productivity. As organizations increasingly ship code t

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation

DGX agent

arXiv:2605.29829v1 Announce Type: new Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems from natural language has emerged as an efficient par

model-releasesarxiv-cs-ai
29 May 2026
Safety

PEARL: Training Socratic Tutors with Pedagogically Aligned Reinforcement Learning

DGX agent

arXiv:2605.29582v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise as educational tutors, yet effective tutoring requires more than solving problems: it must provide pro

safetyarxiv-cs-cl
29 May 2026
Safety

REST3D: Reconstructing Physically Stable 3D Scenes from a Single Image

DGX agent

arXiv:2605.30338v1 Announce Type: new Abstract: Reconstructing physically stable 3D scenes from a single RGB image enables casual images to be converted into simulation-ready digital assets for applic

safetyarxiv-cs-cv
29 May 2026
Research

SalsaAgent: A multimodal embodied language model for interactive dance generation

DGX agent

arXiv:2605.29219v1 Announce Type: new Abstract: Interaction between humanoids involves bidirectional and nonverbal reactivity, coordination and synchrony. Toward socially aware robots and interactive

researcharxiv-cs-cv
29 May 2026
Model Releases

The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More

DGX agent

arXiv:2603.23971v2 Announce Type: replace-cross Abstract: Developers and consumers increasingly choose reasoning models (RMs) based on their listed API prices. However, how accurately do these prices

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

DGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

model-releasesarxiv-cs-lg
29 May 2026
Research

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

DGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

researcharxiv-cs-cl
28 May 2026
Safety

Commit to the Bit: Reactive Reinforcement Learning Done Right

DGX agent

arXiv:2605.28276v1 Announce Type: new Abstract: Reinforcement learning algorithms are commonly analyzed (and designed) under the Markov assumption. This is unrealistic, as most environments encountere

safetyarxiv-cs-lg
28 May 2026
Model Releases

Debate with Images: Detecting Deceptive Behaviors in Multimodal Large Language Models

DGX agent

arXiv:2512.00349v2 Announce Type: replace Abstract: Are frontier AI systems becoming more capable? Certainly. Yet such progress is not an unalloyed blessing but rather a Trojan horse: behind their per

model-releasesarxiv-cs-ai
28 May 2026
Safety

Delay-Aware Reinforcement Learning for Highway On-Ramp Merging under Stochastic Communication Latency

DGX agent

arXiv:2403.11852v5 Announce Type: replace-cross Abstract: Delayed and partially observable state information poses significant challenges for reinforcement learning (RL)-based control in real-world au

safetyarxiv-cs-ai
28 May 2026
Model Releases

DisasterBench: Benchmarking LLM Planning under Typed Tool Interface Constraints

DGX agent

arXiv:2605.27957v1 Announce Type: new Abstract: Disasters cause severe societal impacts, demanding rapid coordination of heterogeneous AI tools, from satellite analysis to flood prediction and damage

model-releasesarxiv-cs-cl
28 May 2026
Research

EchoAvatar: Real-time Generative Avatar Animation from Audio Streams

DGX agent

arXiv:2605.28272v1 Announce Type: new Abstract: Real-time synthesis of high-fidelity 3D character motion from audio is a pivotal component for next-generation interactive avatars and virtual assistant

researcharxiv-cs-cv
28 May 2026
Model Releases

ESC-Skills: Discovering and Self-Evolving Skills for Emotional Support Conversations

DGX agent

arXiv:2605.27908v1 Announce Type: cross Abstract: Existing emotional support conversation (ESC) systems mainly rely on end-to-end response generation or coarse strategy supervision, offering limited i

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

How Far Can Disaggregation Go? A Design-Space Exploration of Attention-FFN Disaggregation for Efficient MoE LLM Serving

DGX agent

arXiv:2605.28302v1 Announce Type: cross Abstract: Modern large language model (LLM) inference has progressively disaggregated to keep pace with growing model sizes and tight TTFT and TPOT service-leve

model-releasesarxiv-cs-ai
28 May 2026
Local Ai

OmniVerifier-M1: Multimodal Meta-Verifier with Explicit Structured Recalibration

DGX agent

arXiv:2605.28805v1 Announce Type: cross Abstract: Visual outcomes are increasingly central to multimodal large language models, making reliable and fine-grained verification essential for scaling gene

local-aiarxiv-cs-ai
28 May 2026
Applications

OphIn-500K: Curating Web-Scale Visual Instructions for Scaling Ophthalmic Multimodal Large Language Models

DGX agent

arXiv:2605.27916v1 Announce Type: cross Abstract: The advancement of general medical Multimodal Large Language Models (MLLMs) has shown great potential for building conversational assistants to suppor

applicationsarxiv-cs-cl
28 May 2026
Model Releases

OralAgent: Integrating Reasoning, Tools, and Knowledge for Interactive Dental Image Analysis

DGX agent

arXiv:2605.27378v1 Announce Type: new Abstract: Dental image analysis plays a pivotal role in supporting accurate diagnosis and treatment planning in oral healthcare. Although recent advances have pro

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

Plug-and-Play Benchmarking of Reinforcement Learning Algorithms for Large-Scale Flow Control

DGX agent

arXiv:2601.15015v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown promising results in active flow control (AFC), yet progress in the field remains difficult to assess as exist

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

POINav: Benchmarking and Enhancing Final-Meters Arrival in Real-World Vision-Language Navigation

DGX agent

arXiv:2605.28237v1 Announce Type: cross Abstract: Real-world navigation is fundamentally driven by Points of Interest (POIs), yet reaching a precise POI remains a critical 'final-meters' challenge. Ex

model-releasesarxiv-cs-cv
28 May 2026
Model Releases

Snowveil: A Framework for Decentralised Preference Discovery

DGX agent

arXiv:2512.18444v2 Announce Type: replace-cross Abstract: Aggregating subjective preferences in social choice traditionally assumes a trusted central authority. In contrast, this paper formalises Dece

model-releasesarxiv-cs-ai
28 May 2026
Safety

Teacher-Student Representational Alignment for Reinforcement Learning-Driven Imitation Learning

DGX agent

arXiv:2605.28372v1 Announce Type: new Abstract: Imitation learning (IL) from a state-based reinforcement learning (RL) policy is a common approach to overcome the curse of dimensionality in complex an

safetyarxiv-cs-lg
28 May 2026
Safety

The Illusion of Opting in AI-Mediated Consequential Decisions

DGX agent

arXiv:2605.28210v1 Announce Type: new Abstract: Drawing on Ullmann-Margalit's concept of opting (transformative, irrevocable, and shadowed by foreclosed alternatives), we show that current AI systems

safetyarxiv-cs-ai
28 May 2026
Safety

VLA-Hijack: A Transferable Patch Attack against Vision-Language-Action Models via Visual Proprioception Hijacking

DGX agent

arXiv:2605.28083v1 Announce Type: new Abstract: While Vision-Language-Action (VLA) models have emerged as powerful generalist policies, their severe vulnerability to adversarial patches significantly

safetyarxiv-cs-cv
28 May 2026
Safety

Alignment Tuning for Large Language Models: A Data-Centric Lens on Alignment Data Pipelines

DGX agent

arXiv:2605.26442v1 Announce Type: cross Abstract: Much of the alignment tuning literature is organized around optimization objectives, while the construction of alignment data is often treated implici

safetyarxiv-cs-ai
27 May 2026
Hardware

AssetGen: Deployable 3D Asset Generation at Interactive Speed

DGX agent

arXiv:2605.26137v1 Announce Type: cross Abstract: While 3D generation is progressing rapidly, recent work has often focused on obtaining high-resolution assets, leaving user experience and deployabili

hardwarearxiv-cs-ai
27 May 2026
Safety

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty

DGX agent

arXiv:2605.26627v1 Announce Type: cross Abstract: Deploying reinforcement learning in safety critical domains, from autonomous vehicles to medical decision support, is constrained by failures arising

safetyarxiv-cs-ro
27 May 2026
Model Releases

ChartAct: A Benchmark for Dynamic Chart Understanding

DGX agent

arXiv:2605.26994v1 Announce Type: new Abstract: Charts are widely used to present complex data for analysis and decision making. Existing chart understanding benchmarks mainly focus on static charts,

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

EgoProx: Evaluating MLLMs on Egocentric 3D Proximity Reasoning Across a Cognitive Hierarchy

DGX agent

arXiv:2605.24456v2 Announce Type: replace Abstract: Humans constantly reason about 3D proximity, the relations between their body and surrounding objects, to guide perception and action in daily life.

model-releasesarxiv-cs-cv
27 May 2026
Applications

Examining the Challenges of Intellectual Property in AI-Generated Productions

DGX agent

arXiv:2605.26590v1 Announce Type: cross Abstract: With the advancement of artificial intelligence systems capable of autonomously generating artistic, literary, musical works, and even inventions with

applicationsarxiv-cs-ai
27 May 2026
Local Ai

Explainable Cross-Disease Reasoning for Cardiovascular Risk Assessment from Low-Dose Computed Tomography

DGX agent

arXiv:2511.06625v5 Announce Type: replace-cross Abstract: Low-dose chest computed tomography (LDCT) captures pulmonary and cardiac structures in a single scan, enabling joint assessment of lung and ca

local-aiarxiv-cs-ai
27 May 2026
Research

Innovative Silicosis and Pneumonia Classification: Leveraging Graph Transformer Post-hoc Modeling and Ensemble Techniques

DGX agent

arXiv:2501.00520v2 Announce Type: replace Abstract: This paper presents a comprehensive study on the classification and detection of Silicosis-related lung inflammation. Our main contributions include

researcharxiv-cs-cv
27 May 2026
Model Releases

Learning to Predict Future-Aligned Research Proposals with Language Models

DGX agent

arXiv:2603.27146v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to assist ideation in research, but evaluating the quality of LLM-generated research proposals re

model-releasesarxiv-cs-cl
27 May 2026
Model Releases

LLM-guided Hierarchical Search for End-to-end Reasoning Intensive Retrieval

DGX agent

arXiv:2510.13217v2 Announce Type: replace-cross Abstract: Search systems are increasingly used for reasoning-intensive queries, where what makes a document relevant requires understanding or reasoning

model-releasesarxiv-cs-lg
27 May 2026
← Previous
1…213214215216217…230
Next →