AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Mega-ASR: Towards In-the-wild^2 Speech Recognition via Scaling up Real-world Acoustic Simulation

DGX agent

arXiv:2605.19833v1 Announce Type: cross Abstract: Despite rapid advances in automatic speech recognition (ASR) and large audio-language models, robust recognition in real-world environments remains li

safetyarxiv-cs-ai
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Memory-Augmented Reinforcement Learning Agent for CAD Generation

DGX agent

arXiv:2605.19748v1 Announce Type: new Abstract: Automatic generation of computer-aided design (CAD) models is a core technology for enabling intelligence in advanced manufacturing. Existing generation

safetyarxiv-cs-ai
20 May 2026
Safety

Metric-Gradient Projection for Stable Multi-Agent Policy Learning

DGX agent

arXiv:2605.18809v1 Announce Type: cross Abstract: General-sum multi-agent learning is often governed by a stacked update field in which each agent's policy update changes the optimization landscape fa

safetyarxiv-cs-ai
20 May 2026
Safety

Multi-Session Ground Texture SLAM in Low-Dynamic Environments

DGX agent

arXiv:2605.19701v1 Announce Type: new Abstract: The simultaneous localization and mapping community has introduced a growing number of systems adapted for multi-session operations where the operationa

safetyarxiv-cs-ro
20 May 2026
Safety

Neuron Incidence Redistribution for Fairness in Medical Image Classification

DGX agent

arXiv:2605.19393v1 Announce Type: new Abstract: Deep learning models for medical image classification are susceptible to subgroup performance disparities across demographic attributes such as age, gen

safetyarxiv-cs-cv
20 May 2026
Safety

Noise-corrected GRPO: From Noisy Rewards to Unbiased Gradients

DGX agent

arXiv:2510.18924v3 Announce Type: replace-cross Abstract: Reinforcement learning from human feedback (RLHF) or verifiable rewards (RLVR), the standard paradigm for aligning LLMs or building recent SOT

safetyarxiv-cs-ai
20 May 2026
Safety

Not all uncertainty is alike: volatility, stochasticity, and exploration

DGX agent

arXiv:2605.19215v1 Announce Type: new Abstract: Adaptive decision-making in biological and artificial intelligence requires balancing the exploitation of known outcomes with the exploration of uncerta

safetyarxiv-cs-ai
20 May 2026
Safety

Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR

DGX agent

arXiv:2605.20164v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards has made post-training highly effective when correctness can be checked automatically. However, many impo

safetyarxiv-cs-ai
20 May 2026
Safety

One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer

DGX agent

arXiv:2511.22940v3 Announce Type: replace Abstract: Recent advances in diffusion models have greatly improved pose-driven character animation. However, existing methods are limited to spatially aligne

safetyarxiv-cs-cv
20 May 2026
Safety

Optimal Representation Size: High-Dimensional Analysis of Pretraining and Linear Probing

DGX agent

arXiv:2605.20105v1 Announce Type: new Abstract: Learning to generalise from limited data is a fundamental challenge for both artificial and biological systems. A common strategy is to extract reusable

safetyarxiv-cs-lg
20 May 2026
Safety

PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models

DGX agent

arXiv:2605.19580v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models show promising ability in language-guided robotic tasks. However, making VLA policies reliable remains challenging,

safetyarxiv-cs-ro
20 May 2026
Safety

PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents

DGX agent

arXiv:2605.19932v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly operate over long and recurring external contexts, like document corpora and code repositories. Across in

safetyarxiv-cs-ai
20 May 2026
Safety

Phase-Aware Mixture of Experts for Agentic Reinforcement Learning

DGX agent

arXiv:2602.17038v3 Announce Type: replace Abstract: Reinforcement learning (RL) has equipped LLM agents with a strong ability to solve complex tasks. However, existing RL methods normally use a single

safetyarxiv-cs-ai
20 May 2026
Safety

Physics-informed simulation framework for realistic sonar image generation and statistical validation

DGX agent

arXiv:2605.19712v1 Announce Type: new Abstract: Synthetic sonar datasets offer a scalable alternative to costly real-world acquisition, yet their utility remains limited by the absence of rigorous qua

safetyarxiv-cs-cv
20 May 2026
Safety

Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance

DGX agent

arXiv:2605.18801v1 Announce Type: new Abstract: Data is fundamental to large language models (LLMs). However, understanding of what makes certain data useful for different stages of an LLM workflow, i

safetyarxiv-cs-ai
20 May 2026
Safety

Precision Physical Activity Prescription via Reinforcement Learning for Functional Actions

DGX agent

arXiv:2605.19208v1 Announce Type: cross Abstract: Physical activity (PA) plays an important role in maintaining and improving health. Daily steps have been a key PA measure that is easily accessible w

safetyarxiv-cs-lg
20 May 2026
Safety

Prediction Is Not Physics: Learning and Evaluating Conserved Quantities in Neural Simulators

DGX agent

arXiv:2605.18883v1 Announce Type: cross Abstract: A diffusion model trained on Hamiltonian trajectories can achieve rollout MSE near 10^{-3}, but the standard deviation of its energy over time is betw

safetyarxiv-cs-ai
20 May 2026
Safety

Probabilistic Multivariate Time Series Forecasting with Diffusion Copulas

DGX agent

arXiv:2605.19685v1 Announce Type: cross Abstract: Accurately assessing financial risk requires capturing both individual asset volatility and the complex, asymmetric dependence structures that emerge

safetyarxiv-cs-lg
20 May 2026
Safety

Progressive Autonomy as Preference Learning: A Formalization of Trust Calibration for Agentic Tool Use

DGX agent

arXiv:2605.19151v1 Announce Type: new Abstract: We formalize trust calibration for agentic tool use (deciding when an automated agent's proposed action may execute autonomously versus require human ap

safetyarxiv-cs-ai
20 May 2026
Safety

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

DGX agent

arXiv:2605.18803v1 Announce Type: cross Abstract: Modern action-conditioned video world models achieve strong short-horizon visual realism, yet remain unreliable on rare, interaction-critical transiti

safetyarxiv-cs-ai
20 May 2026
Safety

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

DGX agent

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

safetyarxiv-cs-ai
20 May 2026
Safety

Rapid patient-specific neural networks for intraoperative X-ray to volume registration

DGX agent

arXiv:2503.16309v2 Announce Type: replace-cross Abstract: Advanced navigation techniques in image-guided interventions and surgical robotics require the rapid and precise alignment of 3D preoperative

safetyarxiv-cs-cv
20 May 2026
Safety

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era

DGX agent

arXiv:2605.18903v1 Announce Type: cross Abstract: Vision-Language Models in Continual Learning (VLM-CL) aim to continuously adapt to new multimodal tasks while retaining prior knowledge. The emerging

safetyarxiv-cs-cv
20 May 2026
Safety

Rethinking the Design Space of Reinforcement Learning for Diffusion Models: On the Importance of Likelihood Estimation Beyond Loss Design

DGX agent

arXiv:2602.04663v2 Announce Type: replace-cross Abstract: Reinforcement learning has been widely applied to diffusion and flow models for visual tasks such as text-to-image generation. However, these

safetyarxiv-cs-ai
20 May 2026
Safety

Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents

DGX agent

arXiv:2605.20061v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) is a promising paradigm for improving large language model (LLM) agents on long-horizon interactiv

safetyarxiv-cs-cl
20 May 2026
Safety

RLFTSim: Realistic and Controllable Multi-Agent Traffic Simulation via Reinforcement Learning Fine-Tuning

DGX agent

arXiv:2605.19033v1 Announce Type: cross Abstract: Supervised open-loop training has been widely adopted for training traffic simulation models; however, it fails to capture the inherently dynamic, mul

safetyarxiv-cs-ai
20 May 2026
Safety

RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields

DGX agent

arXiv:2412.02818v4 Announce Type: replace-cross Abstract: Robot manipulation policies, while central to the promise of physical AI, are highly vulnerable in the presence of external variations in the

safetyarxiv-cs-lg
20 May 2026
Safety

RoHIL: Robust Human-in-the-Loop Robotic Reinforcement Learning Against Illumination Variations

DGX agent

arXiv:2605.19924v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning systems achieve near-perfect success on the workstation where they are trained, but collapse when the same robo

safetyarxiv-cs-ro
20 May 2026
Safety

RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models

DGX agent

arXiv:2605.19678v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance on embodied manipulation, yet they remain brittle under visual observation changes, pa

safetyarxiv-cs-ro
20 May 2026
Safety

SAGE: Scalable Automatic Gating Ensemble for Confident Negative Harvesting in Fraud Detection

DGX agent

arXiv:2605.20157v1 Announce Type: new Abstract: Music streaming fraud, where bad actors artificially inflate stream counts to manipulate chart rankings and royalty payments, poses a significant threat

safetyarxiv-cs-lg
20 May 2026
Safety

SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs

DGX agent

arXiv:2605.18864v1 Announce Type: cross Abstract: Recent studies observe that reinforcement learning with verifiable rewards (RLVR) reliably improves pass@1 on reasoning tasks, yet often fails to yiel

safetyarxiv-cs-ai
20 May 2026
Safety

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling

DGX agent

arXiv:2503.06310v4 Announce Type: replace Abstract: Generating coherent long-form video sequences from discrete text prompts remains challenging due to difficulties in maintaining temporal coherence,

safetyarxiv-cs-cv
20 May 2026
Safety

SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects

DGX agent

arXiv:2605.19587v1 Announce Type: new Abstract: Indoor scene synthesis underpins embodied AI, robotic manipulation, and simulation-based policy evaluation, where a useful scene must specify not only w

safetyarxiv-cs-ai
20 May 2026
Safety

Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting

DGX agent

arXiv:2605.19554v1 Announce Type: new Abstract: Instilling creativity in text-to-image (T2I) generation presents a significant challenge, as it requires synthesized images to exhibit not only visual n

safetyarxiv-cs-cv
20 May 2026
Safety

Set-Valued Policy Learning

DGX agent

arXiv:2605.19830v1 Announce Type: new Abstract: Conventional treatment policies map patient covariates to a single recommended intervention in order to maximize expected clinical outcomes. Although a

safetyarxiv-cs-lg
20 May 2026
Safety

SimGym: A Framework for A/B Test Simulation in E-Commerce with Traffic-Grounded VLM Agents

DGX agent

arXiv:2605.19219v1 Announce Type: new Abstract: A/B testing remains the gold standard for evaluating modifications to e-commerce storefronts, yet it diverts traffic, requires weeks to reach statistica

safetyarxiv-cs-ai
20 May 2026
Safety

Smooth Partial Lotteries for Stable Randomized Selection

DGX agent

arXiv:2605.20069v1 Announce Type: new Abstract: Competitive selection processes, from scientific funding to admissions and hiring, use evaluations to score candidates, and eventually choose a subset o

safetyarxiv-cs-lg
20 May 2026
Safety

Spatially Prompted Visual Trajectory Prediction for Egocentric Manipulation

DGX agent

arXiv:2605.20085v1 Announce Type: new Abstract: Robotic manipulation is often specified through language instructions or task identifiers, yet cluttered environments with similar objects are better ha

safetyarxiv-cs-cv
20 May 2026
Safety

Stitched Value Model for Diffusion Alignment

DGX agent

arXiv:2605.19804v1 Announce Type: cross Abstract: For practical use, diffusion- or flow-based generative models must be aligned with task-specific rewards, such as prompt fidelity or aesthetic prefere

safetyarxiv-cs-ai
20 May 2026
Safety

Structural Energy Guidance for View-Consistent Text-to-3D Generation

DGX agent

arXiv:2605.19876v1 Announce Type: new Abstract: Text-to-3D generation based on diffusion models often suffers from the Janus problem, leading to inconsistent geometry across viewpoints. This work iden

safetyarxiv-cs-cv
20 May 2026
Safety

StruMPL: Multi-task Dense Regression under Disjoint Partial Supervision and MNAR Labels

DGX agent

arXiv:2605.19931v1 Announce Type: cross Abstract: Estimating forest aboveground biomass (AGB) from Earth observation combines two structurally incompatible label sources: spaceborne lidar provides can

safetyarxiv-cs-ai
20 May 2026
Safety

Swimming with Whales: Analysis of Power Imbalances in Stake-Weighted Governance

DGX agent

arXiv:2605.19264v1 Announce Type: new Abstract: Voting methods weighted by stakes are the fundamental governance paradigm in Proof-of-Stake (PoS) blockchains. Such a paradigm is known to be prone to p

safetyarxiv-cs-ai
20 May 2026
Safety

Symmetry in the Wild: The Role of Equivariance in Neural Fluid Surrogates

DGX agent

arXiv:2605.18816v1 Announce Type: cross Abstract: Neural surrogates enable orders-of-magnitude acceleration of computational fluid dynamics (CFD) simulations, with the potential to transform engineeri

safetyarxiv-cs-ai
20 May 2026
Safety

TEA-Time: Transporting Effects Across Time

DGX agent

arXiv:2603.07018v2 Announce Type: replace-cross Abstract: Treatment effects estimated from a randomized controlled trial are local not only to the study population but also to the time at which the tr

safetyarxiv-cs-lg
20 May 2026
Safety

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

DGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

safetyarxiv-cs-lg
20 May 2026
Safety

Text-to-SPARQL Generation with Reinforcement Learning: A GRPO-based Approach on DBLP

DGX agent

arXiv:2605.20066v1 Announce Type: new Abstract: Knowledge graph question answering seeks to translate natural language questions into executable queries over knowledge graphs, but existing approaches

safetyarxiv-cs-cl
20 May 2026
Safety

The Accessibility Capability Boundary: Operational Limits and Expansion Potential of AI-Generated Browser-Native Accessibility Systems

DGX agent

arXiv:2605.19638v1 Announce Type: cross Abstract: As large language models (LLMs) demonstrate increasing competence in synthesizing functional user interfaces, a fundamental question emerges in access

safetyarxiv-cs-ai
20 May 2026
Safety

ThoughtTrace: Understanding User Thoughts in Real-World LLM Interactions

DGX agent

arXiv:2605.20087v1 Announce Type: cross Abstract: Conversational AI has now reached billions of users, yet existing datasets capture only what people say, not what they think. We introduce ThoughtTrac

safetyarxiv-cs-ai
20 May 2026
← Previous
1…170171172173174…260
Next →