AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
Human
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
61,904 results
20 May 2026

Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees

ResearchDGX agent

arXiv:2605.18775v1 Announce Type: cross Abstract: Graph-based Retrieval-Augmented Generation (RAG) systems leverage interconnected knowledge structures to capture complex relationships that flat retri

Query-Conditioned Graph Retrieval for Contextualized LLM Reasoning in Personalized Wearable Data

Local AiDGX agent

arXiv:2605.18763v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly applied to analyzing wearable sensing data, which are long-term, multimodal, and highly personalized. A

Rapid patient-specific neural networks for intraoperative X-ray to volume registration

SafetyDGX agent

arXiv:2503.16309v2 Announce Type: replace-cross Abstract: Advanced navigation techniques in image-guided interventions and surgical robotics require the rapid and precise alignment of 3D preoperative

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RE-VLM: Event-Augmented Vision-Language Model for Scene Understanding

Model ReleasesDGX agent

arXiv:2605.19329v1 Announce Type: cross Abstract: Conventional vision-language models (VLMs) struggle to interpret scenes captured under adverse conditions (e.g., low light, high dynamic range, or fas

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

Model ReleasesDGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

Real-Time Parallel Counterfactual Regret Minimization

Local AiDGX agent

arXiv:2605.19928v1 Announce Type: cross Abstract: Counterfactual Regret Minimization (CFR) is the dominant algorithmic family for solving large imperfect-information games, underpinning breakthroughs

Real-World On-Vehicle Evaluation of Embedding-Based Anomaly Detection

Model ReleasesDGX agent

arXiv:2605.19744v1 Announce Type: new Abstract: Detecting anomalies in traffic scenes is crucial for ensuring safety in autonomous driving, yet collecting representative anomalous data remains challen

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era

SafetyDGX agent

arXiv:2605.18903v1 Announce Type: cross Abstract: Vision-Language Models in Continual Learning (VLM-CL) aim to continuously adapt to new multimodal tasks while retaining prior knowledge. The emerging

Rebalancing Reference Frame Dominance to Improve Motion in Image-to-Video Models

Model ReleasesDGX agent

arXiv:2605.19398v1 Announce Type: cross Abstract: Image-to-video models often generate videos that remain overly static, compared to text-to-video models. While prior approaches mitigate this issue by

Receptogenesis in a Vascularized Robotic Embodiment

ResearchDGX agent

arXiv:2603.09473v2 Announce Type: replace Abstract: Equipping robotic systems with the capacity to generate extit{ex novo} hardware during operation extends control of physical adaptability. Unlike mo

RECIPE: Procedural Planning via Grounding in Instructional Video

Model ReleasesDGX agent

arXiv:2605.19976v1 Announce Type: new Abstract: Visual planning asks a model to generate the remaining steps of a procedure in natural language given a partial video context and a goal. Progress on th

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents

Model ReleasesDGX agent

arXiv:2605.18805v1 Announce Type: cross Abstract: LLM recommendation agents increasingly produce structured recommendation reports: sets of items accompanied by natural-language justifications. Yet ex

ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning

ResearchDGX agent

arXiv:2605.18799v1 Announce Type: cross Abstract: Large language models can fail in critic interaction not only by answering incorrectly, but also by abandoning an initially correct scientific solutio

Recursive Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model

Model ReleasesDGX agent

arXiv:2506.00286v3 Announce Type: replace-cross Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs with recursive entropic risk measures (ERM), where the risk parameter

Reducing Diffusion Model Memorization with Higher Order Langevin Dynamics

ApplicationsDGX agent

arXiv:2605.19170v1 Announce Type: cross Abstract: Diffusion/score-based models have emerged as powerful generative models, capable of generating high-quality samples that mimic the training data distr

Reflection-Based Relative Localization for Cooperative UAV Teams Using Active Markers

ResearchDGX agent

arXiv:2511.17166v2 Announce Type: replace Abstract: Reflections of active markers in the environment are a common source of ambiguity in onboard visual relative localization. This work presents a nove

Replacement Learning: Training Neural Networks with Fewer Parameters

Model ReleasesDGX agent

arXiv:2605.19533v1 Announce Type: new Abstract: End-to-end training with full-depth backpropagation remains the dominant paradigm for optimizing deep neural networks, but its efficiency deteriorates a

Resilient Byzantine Agreement with Predictions

Model ReleasesDGX agent

arXiv:2605.19452v1 Announce Type: cross Abstract: This paper studies the Byzantine Agreement problem where the nodes have access to a predictor that flags nodes for suspicion of faulty (Byzantine) beh

Rethinking How to Remember: Beyond Atomic Facts in Lifelong LLM Agent Memory

Model ReleasesDGX agent

arXiv:2605.19952v1 Announce Type: new Abstract: To enable reliable long-term interaction, LLM agents require a memory system that can faithfully store, efficiently retrieve, and deeply reason over acc

Rethinking Muon Beyond Pretraining: Spectral Failures and High-Pass Remedies for VLA and RLVR

ResearchDGX agent

arXiv:2605.19282v1 Announce Type: new Abstract: Muon is a matrix-aware optimizer that leverages Newton-Schulz (NS) iterations to enforce spectral gradient orthogonalization by driving all singular val

Rethinking the Design Space of Reinforcement Learning for Diffusion Models: On the Importance of Likelihood Estimation Beyond Loss Design

SafetyDGX agent

arXiv:2602.04663v2 Announce Type: replace-cross Abstract: Reinforcement learning has been widely applied to diffusion and flow models for visual tasks such as text-to-image generation. However, these

Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models

Local AiDGX agent

arXiv:2605.20158v1 Announce Type: cross Abstract: Large Vision Language Models (LVLMs) show promise in medical applications, but their inability to faithfully ground responses in visual evidence raise

Retrieval-Augmented Generation for Natural Language Processing: A Survey

Model ReleasesDGX agent

arXiv:2407.13193v4 Announce Type: replace Abstract: Large language models (LLMs) have achieved strong empirical performance in various fields, benefiting from their huge amount of parameters that stor

Retrieval-Augmented Linguistic Calibration

ResearchDGX agent

arXiv:2605.19344v1 Announce Type: new Abstract: Linguistic cues such as 'I believe' and 'probably' offer an intuitive interface for communicating confidence, yet a generalisable, principled calibratio

Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method

ResearchDGX agent

arXiv:2605.18766v1 Announce Type: cross Abstract: Retrieving relevant tables from extensive databases for a given natural language query is essential for accurately answering questions in tasks such a

Return of Frustratingly Easy Unsupervised Video Domain Adaptation

ResearchDGX agent

arXiv:2605.19510v1 Announce Type: new Abstract: Unsupervised video domain adaptation (UVDA) is a practical but under-explored problem. In this paper, we propose a frustratingly easy UVDA method, calle

Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents

SafetyDGX agent

arXiv:2605.20061v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) is a promising paradigm for improving large language model (LLM) agents on long-horizon interactiv

Rewriting History: A Recipe for Interventional Analyses to Study Data Effects on Model Behavior

Model ReleasesDGX agent

arXiv:2510.14261v2 Announce Type: replace Abstract: We present an experimental recipe for studying the relationship between training data and language model (LM) behavior. We outline steps for interve

Riemannian Networks over Full-Rank Correlation Matrices

ResearchDGX agent

arXiv:2605.19073v1 Announce Type: cross Abstract: Representations on the Symmetric Positive Definite (SPD) manifold have garnered significant attention across different applications. In contrast, the

RLFTSim: Realistic and Controllable Multi-Agent Traffic Simulation via Reinforcement Learning Fine-Tuning

SafetyDGX agent

arXiv:2605.19033v1 Announce Type: cross Abstract: Supervised open-loop training has been widely adopted for training traffic simulation models; however, it fails to capture the inherently dynamic, mul

RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents

Model ReleasesDGX agent

arXiv:2605.19328v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) facilitate a new class of embodied AI systems, where these models are integrated into physical platfo

RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields

SafetyDGX agent

arXiv:2412.02818v4 Announce Type: replace-cross Abstract: Robot manipulation policies, while central to the promise of physical AI, are highly vulnerable in the presence of external variations in the

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

SafetyDGX agent

arXiv:2605.19940v1 Announce Type: new Abstract: Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cum

Robots that learn to evaluate models of collective behavior

Model ReleasesDGX agent

arXiv:2604.07303v2 Announce Type: replace Abstract: Understanding and modeling animal behavior is essential for studying collective motion, decision-making, and bio-inspired robotics. Yet, evaluating

Robust Basis Spline Decoupling for the Compression of Transformer Models

Model ReleasesDGX agent

arXiv:2605.18794v1 Announce Type: cross Abstract: Decoupling is a powerful modeling paradigm for representing multivariate functions as compositions of linear transformations and univariate nonlinear

Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking

AgentsDGX agent

arXiv:2605.18852v1 Announce Type: cross Abstract: Checkpoint selection for multimodal large language models (MLLMs) presents significant challenges when performance differentials are marginal and eval

Robust Mitigation of Age-Dependent Confounding Effects via Sample-Difficulty Decorrelation

ResearchDGX agent

arXiv:2605.19230v1 Announce Type: new Abstract: Age dependent performance disparities in medical image classification often arise because age acts as a confounder, linking imaging morphology with dise

Robustness and Regularization in Hierarchical Re-Basin

ResearchDGX agent

arXiv:2510.09174v3 Announce Type: replace Abstract: This paper takes a closer look at Git Re-Basin, an interesting new approach to merge trained models. We propose a hierarchical model merging scheme

RoHIL: Robust Human-in-the-Loop Robotic Reinforcement Learning Against Illumination Variations

SafetyDGX agent

arXiv:2605.19924v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning systems achieve near-perfect success on the workstation where they are trained, but collapse when the same robo

RoomPilot: Controllable Indoor Scene Synthesis via Multimodal Semantic Parsing

ResearchDGX agent

arXiv:2512.11234v2 Announce Type: replace Abstract: Generating controllable indoor scenes is fundamental to applications in game development, architectural visualization, and embodied AI. However, exi

Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference

ResearchDGX agent

arXiv:2605.19218v1 Announce Type: cross Abstract: Vision-Language Models suffer severe KV cache pressure at inference, as a single image often encodes into thousands of tokens. Most existing methods e

RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.19678v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance on embodied manipulation, yet they remain brittle under visual observation changes, pa

Safe Continual Reinforcement Learning under Nonstationarity via Adaptive Safety Constraints

SafetyDGX agent

arXiv:2605.18842v1 Announce Type: new Abstract: Safe reinforcement learning in nonstationary environments requires safety mechanisms that adapt as environmental conditions change. Standard safe reinfo

SafeAlign-VLA: A Negative-Enhanced Safe Alignment Framework for Risk-Aware Autonomous Driving

SafetyDGX agent

arXiv:2605.19524v1 Announce Type: cross Abstract: End-to-end autonomous driving systems excel in common scenarios but struggle with safety-critical long-tail cases. Vision-Language-Action (VLA) models

SAGA: A Sequence-Adaptive Generative Architecture for Multi-Horizon Probabilistic Forecasting with Adaptive Temporal Conformal Prediction

Model ReleasesDGX agent

arXiv:2605.19014v1 Announce Type: new Abstract: Microsimulation models used by ministries of finance and central banks rely on parametric processes for lifetime earnings that capture only first and se

SAGE: Scalable Automatic Gating Ensemble for Confident Negative Harvesting in Fraud Detection

SafetyDGX agent

arXiv:2605.20157v1 Announce Type: new Abstract: Music streaming fraud, where bad actors artificially inflate stream counts to manipulate chart rankings and royalty payments, poses a significant threat

SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs

SafetyDGX agent

arXiv:2605.18864v1 Announce Type: cross Abstract: Recent studies observe that reinforcement learning with verifiable rewards (RLVR) reliably improves pass@1 on reasoning tasks, yet often fails to yiel

Sampling-Based Safe Reinforcement Learning

SafetyDGX agent

arXiv:2605.19469v1 Announce Type: cross Abstract: Safe exploration remains a fundamental challenge in reinforcement learning (RL), limiting the deployment of RL agents in the real world. We propose Sa

SCAFDS: Edge-Feature Graph Attention for Interbank Fraud Detection with Attribution-Grounded SAR Generation

ResearchDGX agent

arXiv:2605.18913v1 Announce Type: cross Abstract: The U.S. financial system processes approximately 1.3 million interbank transactions daily, yet no system in the reviewed literature models fraud prop

Scalable, Energy-Efficient Optical-Neural Architecture for Multiplexed Deepfake Video Detection

ApplicationsDGX agent

arXiv:2605.19360v1 Announce Type: new Abstract: The rapid proliferation of AI-generated visual media has created an urgent need for efficient, trustworthy deepfake detection systems. However, existing

Scaling Evaluation-time Compute with Reasoning Models as Evaluators

ResearchDGX agent

arXiv:2503.19877v2 Announce Type: replace Abstract: As language model (LM) outputs get more and more natural, it is becoming more difficult than ever to evaluate their quality. Simultaneously, increas

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling

SafetyDGX agent

arXiv:2503.06310v4 Announce Type: replace Abstract: Generating coherent long-form video sequences from discrete text prompts remains challenging due to difficulties in maintaining temporal coherence,

SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects

SafetyDGX agent

arXiv:2605.19587v1 Announce Type: new Abstract: Indoor scene synthesis underpins embodied AI, robotic manipulation, and simulation-based policy evaluation, where a useful scene must specify not only w

ScheduleFree+: Scaling Learning-Rate-Free & Schedule-Free Learning to Large Language Models

Model ReleasesDGX agent

arXiv:2605.19095v1 Announce Type: cross Abstract: Schedule-Free Learning has shown promise as a practical anytime training method for machine learning, showing success across dozens of standard benchm

SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models

Model ReleasesDGX agent

arXiv:2605.19357v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied to scientific research, yet existing evaluations often fail to reflect the fine-grained capabiliti

SEAL: Semantic Aware Image Watermarking

ResearchDGX agent

arXiv:2503.12172v4 Announce Type: replace-cross Abstract: Generative models have rapidly evolved to generate realistic outputs. However, their synthetic outputs increasingly challenge the clear distin

Search Self-play: Pushing the Frontier of Agent Capability without Supervision

Model ReleasesDGX agent

arXiv:2510.18821v3 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become the mainstream technique for training LLM agents. However, RLVR highly depends on w

Selective, Regularized, and Calibrated: Harnessing Vision Foundation Models for Cross-Domain Few-Shot Semantic Segmentation

ResearchDGX agent

arXiv:2605.19340v1 Announce Type: new Abstract: Vision foundation models (VFMs) have achieved strong performance across various vision tasks. However, it still remains challenging to apply VFMs for cr

Self-assembling Modular Aerial Robot for Versatile Aerial Tasks

ResearchDGX agent

arXiv:2605.19431v1 Announce Type: new Abstract: Multirotor aerial robots excel at maneuvering in three-dimensional space, and recent advances enable nimble navigation in cluttered and confined environ

Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting

SafetyDGX agent

arXiv:2605.19554v1 Announce Type: new Abstract: Instilling creativity in text-to-image (T2I) generation presents a significant challenge, as it requires synthesized images to exhibit not only visual n

← Previous
1…645646647648649…1032
Next →