AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
20 May 2026

Position: Let's Develop Data Probes to Fundamentally Understand How Data Affects LLM Performance

SafetyDGX agent

arXiv:2605.18801v1 Announce Type: new Abstract: Data is fundamental to large language models (LLMs). However, understanding of what makes certain data useful for different stages of an LLM workflow, i

Precision Physical Activity Prescription via Reinforcement Learning for Functional Actions

SafetyDGX agent

arXiv:2605.19208v1 Announce Type: cross Abstract: Physical activity (PA) plays an important role in maintaining and improving health. Daily steps have been a key PA measure that is easily accessible w

Prediction Is Not Physics: Learning and Evaluating Conserved Quantities in Neural Simulators

SafetyDGX agent

arXiv:2605.18883v1 Announce Type: cross Abstract: A diffusion model trained on Hamiltonian trajectories can achieve rollout MSE near 10^{-3}, but the standard deviation of its energy over time is betw

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Probabilistic Multivariate Time Series Forecasting with Diffusion Copulas

SafetyDGX agent

arXiv:2605.19685v1 Announce Type: cross Abstract: Accurately assessing financial risk requires capturing both individual asset volatility and the complex, asymmetric dependence structures that emerge

Progressive Autonomy as Preference Learning: A Formalization of Trust Calibration for Agentic Tool Use

SafetyDGX agent

arXiv:2605.19151v1 Announce Type: new Abstract: We formalize trust calibration for agentic tool use (deciding when an automated agent's proposed action may execute autonomously versus require human ap

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

SafetyDGX agent

arXiv:2605.18803v1 Announce Type: cross Abstract: Modern action-conditioned video world models achieve strong short-horizon visual realism, yet remain unreliable on rare, interaction-critical transiti

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

SafetyDGX agent

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

Rapid patient-specific neural networks for intraoperative X-ray to volume registration

SafetyDGX agent

arXiv:2503.16309v2 Announce Type: replace-cross Abstract: Advanced navigation techniques in image-guided interventions and surgical robotics require the rapid and precise alignment of 3D preoperative

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era

SafetyDGX agent

arXiv:2605.18903v1 Announce Type: cross Abstract: Vision-Language Models in Continual Learning (VLM-CL) aim to continuously adapt to new multimodal tasks while retaining prior knowledge. The emerging

Rethinking the Design Space of Reinforcement Learning for Diffusion Models: On the Importance of Likelihood Estimation Beyond Loss Design

SafetyDGX agent

arXiv:2602.04663v2 Announce Type: replace-cross Abstract: Reinforcement learning has been widely applied to diffusion and flow models for visual tasks such as text-to-image generation. However, these

Rewarding Beliefs, Not Actions: Consistency-Guided Credit Assignment for Long-Horizon Agents

SafetyDGX agent

arXiv:2605.20061v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) is a promising paradigm for improving large language model (LLM) agents on long-horizon interactiv

RLFTSim: Realistic and Controllable Multi-Agent Traffic Simulation via Reinforcement Learning Fine-Tuning

SafetyDGX agent

arXiv:2605.19033v1 Announce Type: cross Abstract: Supervised open-loop training has been widely adopted for training traffic simulation models; however, it fails to capture the inherently dynamic, mul

RoboMD: Uncovering Robot Vulnerabilities through Semantic Potential Fields

SafetyDGX agent

arXiv:2412.02818v4 Announce Type: replace-cross Abstract: Robot manipulation policies, while central to the promise of physical AI, are highly vulnerable in the presence of external variations in the

RoHIL: Robust Human-in-the-Loop Robotic Reinforcement Learning Against Illumination Variations

SafetyDGX agent

arXiv:2605.19924v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning systems achieve near-perfect success on the workstation where they are trained, but collapse when the same robo

RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.19678v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance on embodied manipulation, yet they remain brittle under visual observation changes, pa

SAGE: Scalable Automatic Gating Ensemble for Confident Negative Harvesting in Fraud Detection

SafetyDGX agent

arXiv:2605.20157v1 Announce Type: new Abstract: Music streaming fraud, where bad actors artificially inflate stream counts to manipulate chart rankings and royalty payments, poses a significant threat

SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs

SafetyDGX agent

arXiv:2605.18864v1 Announce Type: cross Abstract: Recent studies observe that reinforcement learning with verifiable rewards (RLVR) reliably improves pass@1 on reasoning tasks, yet often fails to yiel

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling

SafetyDGX agent

arXiv:2503.06310v4 Announce Type: replace Abstract: Generating coherent long-form video sequences from discrete text prompts remains challenging due to difficulties in maintaining temporal coherence,

SceneCode: Executable World Programs for Editable Indoor Scenes with Articulated Objects

SafetyDGX agent

arXiv:2605.19587v1 Announce Type: new Abstract: Indoor scene synthesis underpins embodied AI, robotic manipulation, and simulation-based policy evaluation, where a useful scene must specify not only w

Self-Creative Text-to-Object Generation using Semantic-Aware Spatial Weighting

SafetyDGX agent

arXiv:2605.19554v1 Announce Type: new Abstract: Instilling creativity in text-to-image (T2I) generation presents a significant challenge, as it requires synthesized images to exhibit not only visual n

Set-Valued Policy Learning

SafetyDGX agent

arXiv:2605.19830v1 Announce Type: new Abstract: Conventional treatment policies map patient covariates to a single recommended intervention in order to maximize expected clinical outcomes. Although a

SimGym: A Framework for A/B Test Simulation in E-Commerce with Traffic-Grounded VLM Agents

SafetyDGX agent

arXiv:2605.19219v1 Announce Type: new Abstract: A/B testing remains the gold standard for evaluating modifications to e-commerce storefronts, yet it diverts traffic, requires weeks to reach statistica

Smooth Partial Lotteries for Stable Randomized Selection

SafetyDGX agent

arXiv:2605.20069v1 Announce Type: new Abstract: Competitive selection processes, from scientific funding to admissions and hiring, use evaluations to score candidates, and eventually choose a subset o

Spatially Prompted Visual Trajectory Prediction for Egocentric Manipulation

SafetyDGX agent

arXiv:2605.20085v1 Announce Type: new Abstract: Robotic manipulation is often specified through language instructions or task identifiers, yet cluttered environments with similar objects are better ha

Stitched Value Model for Diffusion Alignment

SafetyDGX agent

arXiv:2605.19804v1 Announce Type: cross Abstract: For practical use, diffusion- or flow-based generative models must be aligned with task-specific rewards, such as prompt fidelity or aesthetic prefere

Structural Energy Guidance for View-Consistent Text-to-3D Generation

SafetyDGX agent

arXiv:2605.19876v1 Announce Type: new Abstract: Text-to-3D generation based on diffusion models often suffers from the Janus problem, leading to inconsistent geometry across viewpoints. This work iden

StruMPL: Multi-task Dense Regression under Disjoint Partial Supervision and MNAR Labels

SafetyDGX agent

arXiv:2605.19931v1 Announce Type: cross Abstract: Estimating forest aboveground biomass (AGB) from Earth observation combines two structurally incompatible label sources: spaceborne lidar provides can

Swimming with Whales: Analysis of Power Imbalances in Stake-Weighted Governance

SafetyDGX agent

arXiv:2605.19264v1 Announce Type: new Abstract: Voting methods weighted by stakes are the fundamental governance paradigm in Proof-of-Stake (PoS) blockchains. Such a paradigm is known to be prone to p

Symmetry in the Wild: The Role of Equivariance in Neural Fluid Surrogates

SafetyDGX agent

arXiv:2605.18816v1 Announce Type: cross Abstract: Neural surrogates enable orders-of-magnitude acceleration of computational fluid dynamics (CFD) simulations, with the potential to transform engineeri

TEA-Time: Transporting Effects Across Time

SafetyDGX agent

arXiv:2603.07018v2 Announce Type: replace-cross Abstract: Treatment effects estimated from a randomized controlled trial are local not only to the study population but also to the time at which the tr

TEMPO: Temporal Enforcement via Mode-Separated Policy Optimization for Trustworthy LLM Backtesting

SafetyDGX agent

arXiv:2605.18843v1 Announce Type: new Abstract: Backtesting large language models on historical events requires reasoning exclusively from information available before a specified cutoff date. Yet mod

Text-to-SPARQL Generation with Reinforcement Learning: A GRPO-based Approach on DBLP

SafetyDGX agent

arXiv:2605.20066v1 Announce Type: new Abstract: Knowledge graph question answering seeks to translate natural language questions into executable queries over knowledge graphs, but existing approaches

The Accessibility Capability Boundary: Operational Limits and Expansion Potential of AI-Generated Browser-Native Accessibility Systems

SafetyDGX agent

arXiv:2605.19638v1 Announce Type: cross Abstract: As large language models (LLMs) demonstrate increasing competence in synthesizing functional user interfaces, a fundamental question emerges in access

The consensus of analysts is that AI capital investments will rise by 20% a year for five years while revenues are expected to grow 15% resu…

SafetyDGX agent

The consensus of analysts is that AI capital investments will rise by 20% a year for five years while revenues are expected to grow 15% resulting in negative returns. 'The AI boom will become a story

the crazy part is that people are “clowning” me without knowing anything about the training or whether anything else other than scaled chang…

SafetyDGX agent

the crazy part is that people are “clowning” me without knowing anything about the training or whether anything else other than scaled changed or how the model does on anything else. (or what it costs

The Uncomfortable Truth About AI “Reasoning” | World Science Festival https://youtu.be/iFYF_e1GSGI?si=pFNjvhp48TMKp-PE via @YouTube @GaryMar…

SafetyDGX agent

This video from the World Science Festival, shared by cognitive scientist Gary Marcus, examines the gap between AI systems' claimed reasoning capabilities and their actual mechanisms, likely arguing t

think you know @garymarcus because read a few of his tweets? try watching this, to see the real deal, and why, for example, the US Senate in…

SafetyDGX agent

think you know @garymarcus because read a few of his tweets? try watching this, to see the real deal, and why, for example, the US Senate invited him to testify. Superb conversation between @bgreene a

ThoughtTrace: Understanding User Thoughts in Real-World LLM Interactions

SafetyDGX agent

arXiv:2605.20087v1 Announce Type: cross Abstract: Conversational AI has now reached billions of users, yet existing datasets capture only what people say, not what they think. We introduce ThoughtTrac

Toward an AI-Powered Computational Testbed for Workforce Policy

SafetyDGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

Towards Distillation Guarantees under Algorithmic Alignment for Combinatorial Optimization

SafetyDGX agent

arXiv:2605.20074v1 Announce Type: new Abstract: Distillation transfers knowledge from a large model trained on broad data to a smaller, more efficient model suitable for deployment. In structured pred

Trustworthy Agent Network: Trust in Agent Networks Must Be Baked In, Not Bolted On

SafetyDGX agent

arXiv:2605.19035v1 Announce Type: new Abstract: The rapid advancement of Large Language Models has given rise to autonomous LLM-based agents capable of complex reasoning and execution. As these agents

TSR: Trajectory-Search Rollouts for Multi-Turn RL of LLM Agents

SafetyDGX agent

arXiv:2602.11767v3 Announce Type: replace Abstract: Advances in large language models (LLMs) are driving a shift toward using reinforcement learning (RL) to train agents from iterative, multi-turn int

two trillion dollars to build “pathologically dishonest” AI

SafetyDGX agent

two trillion dollars to build “pathologically dishonest” AI I think @RyanPGreenblatt's recent post summarizes the vibe of this behavior well. Not the sort of thing that would be acceptable for a human

Universal Skeleton Understanding via Differentiable Rendering and MLLMs

SafetyDGX agent

arXiv:2603.18003v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) exhibit strong visual-language reasoning, yet cannot process structured, non-visual data such as human skel

When Critics Disagree: Adaptive Reward Poisoning Attacks in RIS-Aided Wireless Control System

SafetyDGX agent

arXiv:2605.20037v1 Announce Type: cross Abstract: Reward-poisoning attacks present a significant risk to learning-based wireless control systems. Given this, we propose a Disagreement-Guided Reward Po

When Preference Labels Fall Short: Aligning Diffusion Models from Real Data

SafetyDGX agent

arXiv:2605.19839v1 Announce Type: new Abstract: Preference alignment aims to guide generative models by learning from comparisons between preferred and non-preferred samples. In practice, most existin

When Tabular Foundation Models Meet Strategic Tabular Data: A Prior Alignment Approach

SafetyDGX agent

arXiv:2605.19662v1 Announce Type: new Abstract: Tabular foundation models based on pretrained prior-data fitted networks~(PFNs) have shown strong generalization on diverse tabular tasks, but they are

When to Stop Reusing: Dynamic Gradient Gating for Sample-Efficient RLVR

SafetyDGX agent

arXiv:2605.19425v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become the dominant paradigm for advanced reasoning in Large Language Models (LLMs), but rol

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making

SafetyDGX agent

arXiv:2602.07008v2 Announce Type: replace Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typical

Worst-Group Equalized Odds Regularization for Multi-Attribute Fair Medical Image Classification

SafetyDGX agent

arXiv:2605.19214v1 Announce Type: cross Abstract: Diagnostic performance in medical AI varies systematically across demographic groups, yet subgroup AUC can mask clinically important disparities. At a

19 May 2026

A Fourier perspective on the learning dynamics of neural networks: from sample complexities to mechanistic insights

SafetyDGX agent

arXiv:2605.16913v1 Announce Type: cross Abstract: Neural networks trained with gradient-based methods exhibit a strong simplicity bias: they learn simpler statistical features of their data before mov

A Simplex Witness Certificate for Constant Collapse in Variational Autoencoders

SafetyDGX agent

arXiv:2605.18224v1 Announce Type: cross Abstract: This note studies exact constant collapse in variational autoencoders, where the encoder mean becomes independent of the input. The goal is to make th

A study done on 361,645 job applications in almost 30 countries over the last 40 years discovered the hiring bias in society is actually aga…

SafetyDGX agent

A large-scale meta-analysis examining over 361,000 job applications across nearly 30 countries spanning 40 years found that hiring discrimination based on protected characteristics (such as race, gend

A Visual Reinforcement Learning-Based Separate Primitive Policy for Peg-in-Hole Tasks

SafetyDGX agent

arXiv:2504.14820v2 Announce Type: replace Abstract: For peg-in-hole tasks, humans rely on binocular visual perception to locate the peg above the hole surface and then proceed with insertion. This pap

Actionable World Representation

SafetyDGX agent

arXiv:2605.18743v1 Announce Type: new Abstract: Inspired by the emergent behaviors in large language models that generalized human intelligence, the research community is pursuing similar emergent cap

Adaptive Control in Autonomous Driving via Real-Time Recurrent RL

SafetyDGX agent

arXiv:2602.02236v4 Announce Type: replace-cross Abstract: We study online fine-tuning of pretrained control policies for autonomous driving using Real-Time Recurrent Reinforcement Learning (RTRRL), a

Adaptive Experimentation for Censored Survival Outcomes

SafetyDGX agent

arXiv:2605.18459v1 Announce Type: new Abstract: Adaptive experimentation enables efficient estimation of causal effects, but existing methods are not designed for survival data with censoring, where e

Adaptive Generate-Rank-Verify: Inference-Time Search with Costly Verification

SafetyDGX agent

arXiv:2605.17609v1 Announce Type: new Abstract: Many inference-time language-model pipelines combine a cheap reward signal with an expensive verifier, such as exact answer checking in mathematical rea

AffordVLA: Injecting Affordance Representations into Vision-Language-Action Models via Implicit Feature Alignment

SafetyDGX agent

arXiv:2605.17517v1 Announce Type: new Abstract: Recent advances in Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation. However, the visual represe

Agent Bazaar: Enabling Economic Alignment in Multi-Agent Marketplaces

SafetyDGX agent

arXiv:2605.17698v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) as autonomous economic agents introduces systemic risks that extend beyond individual capability failures

← Previous
1…157158159160161…242
Next →