AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
22 May 2026

Governance by Design: Architecting Agentic AI for Organizational Learning and Scalable Autonomy

SafetyDGX agent

arXiv:2605.20210v1 Announce Type: cross Abstract: Agentic AI systems - systems that can pursue goals through multi-step planning and tool-mediated action with limited direct supervision - are moving f

How can reasoning capability empower the AI copilot robot in endoscopic surgery

SafetyDGX agent

arXiv:2605.22322v1 Announce Type: new Abstract: Reasoning capability has significantly advanced complex logical inference and robotic decision-making in general domains. However, its potential in the

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

SafetyDGX agent

arXiv:2605.21401v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that make sequences of decisions over extended interactions in high-stakes

Parallel OctoMapping: A Scalable Framework for Enhanced Path Planning in Autonomous Navigation

SafetyDGX agent

arXiv:2603.22508v2 Announce Type: replace Abstract: Mapping is essential in robotics and autonomous systems because it provides the spatial foundation for path planning. Efficient mapping enables plan

Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO (Washington Post)

SafetyDGX agent

Washington Post: Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO — Industry leaders warned

update @thestalwart found more recent models less vulnerable. would be good to do a broad study of this.

SafetyDGX agent

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models woul

21 May 2026

AI-Assisted Competency Assessment from Egocentric Video in Simulation-Based Nursing Education

SafetyDGX agent

arXiv:2605.20233v1 Announce Type: new Abstract: Assessing learner competency in clinical simulation requires expert observation that is time-intensive, difficult to scale, and subject to inter-rater v

AIMBio-Mat: An AI-Native FAIR Platform for Closed-Loop Materials Discovery and Biomedical Translation

SafetyDGX agent

arXiv:2605.21083v1 Announce Type: cross Abstract: Materials discovery and biomedical translation increasingly require models that can reason across composition, processing, structure, biological respo

CHEM: Estimating and Understanding Hallucinations in Deep Learning for Image Processing

SafetyDGX agent

arXiv:2512.09806v2 Announce Type: replace Abstract: Deep learning-based methods have recently achieved significant success in image reconstruction problems. However, challenges have emerged, as these

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

SafetyDGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

FineVision: Open Data Is All You Need

SafetyDGX agent

arXiv:2510.17269v2 Announce Type: replace Abstract: The advancement of vision-language models (VLMs) is hampered by a fragmented landscape of inconsistent and contaminated public datasets. We introduc

GAMR: Geometric-Aware Manifold Regularization with Virtual Outlier Synthesis for Learning with Noisy Labels

SafetyDGX agent

arXiv:2605.20727v1 Announce Type: new Abstract: Deep neural networks (DNNs) experience significant performance degradation when processing noisy labels, primarily due to overfitting on mislabeled data

GraphDiffMed: Knowledge-Constrained Differential Attention with Pharmacological Graph Priors for Medication Recommendation

SafetyDGX agent

arXiv:2605.20188v1 Announce Type: new Abstract: Recommending safe and effective medication combinations from electronic health records (EHRs) is a core clinical AI problem, yet it remains difficult be

Lost in Fog: Sensor Perturbations Expose Reasoning Fragility in Driving VLAs

SafetyDGX agent

arXiv:2605.21446v1 Announce Type: new Abstract: Interpretable autonomous driving planners depend not only on generating explanations, but also on those explanations remaining reliable under real-world

Mechanistic Interpretability for Learning Assurance of a Vision-Based Landing System

SafetyDGX agent

arXiv:2605.20607v1 Announce Type: cross Abstract: EASA's learning-assurance guidance requires data-driven aviation systems to build and monitor their own situation representation, yet for neural netwo

Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty

SafetyDGX agent

arXiv:2605.20255v1 Announce Type: new Abstract: Simulation-based testing of self-driving cars (SDCs) typically relies on scripted or simplified pedestrian models that do not capture the heterogeneity

Preference-aware Influence-function-based Data Selection Method for Efficient Fine-Tuning

SafetyDGX agent

arXiv:2605.21422v1 Announce Type: new Abstract: As LLMs continue to scale, improving training efficiency increasingly depends on using data more effectively. Data selection addresses this problem by a

Quadratic Characterizations for Reachability Analysis of Neural Networks

SafetyDGX agent

arXiv:2605.20482v1 Announce Type: new Abstract: Quadratic constraints (QCs) are widely used to characterize nonlinearities and uncertainties, but generic analytical characterizations can be conservati

Spacetime Optimal-Transport Attention for Visuo-Haptic Imitation Learning of Contact-Rich Manipulation

SafetyDGX agent

arXiv:2605.20433v1 Announce Type: new Abstract: Contact-rich manipulation tasks such as tight-clearance insertion, connector mating, polishing, and surface-conforming wiping remain difficult for data-

SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary

SafetyDGX agent

arXiv:2605.21132v1 Announce Type: new Abstract: Understanding surgical workflow in real time is fundamental for intelligent surgical embodiment, where AI systems continuously perceive and respond as s

Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs

SafetyDGX agent

arXiv:2605.20641v1 Announce Type: cross Abstract: Inference optimization is a vital technique for deploying LLMs at scale. Compilation is the most widely adopted optimization technique for LLMs. While

When AI Gets it Wrong: Reliability and Risk in AI-Assisted Medication Decision Systems

SafetyDGX agent

arXiv:2604.01449v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) systems are increasingly integrated into healthcare and pharmacy workflows, supporting tasks such as medication r

20 May 2026

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders

SafetyDGX agent

arXiv:2605.19503v1 Announce Type: cross Abstract: Reinforcement learning for legged locomotion has matured into a stack of multi-component reward functions and physics-engine benchmarks whose morpholo

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models

SafetyDGX agent

arXiv:2605.19485v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated remarkable capabilities in solving complex problems by generating structured, step-by-step reasoning con

Backtracking When It Strays: Mitigating Dual Exposure Biases in LLM Reasoning Distillation

SafetyDGX agent

arXiv:2605.19433v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable success in complex reasoning tasks via long chain-of-thought (CoT), yet their immense computatio

Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks

SafetyDGX agent

arXiv:2605.19147v1 Announce Type: cross Abstract: Large language models (LLMs) are highly susceptible to backdoor attacks (BAs), wherein training samples are poisoned using trigger-based harmful conte

Beyond Imitation: Learning Safe End-to-End Autonomous Driving from Hard Negatives

Model ReleasesDGX agent

arXiv:2605.19771v1 Announce Type: cross Abstract: Existing imitation learning methods for end-to-end autonomous driving predominantly learn from successful demonstrations by minimizing geometric devia

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced wit…

SafetyDGX agent

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced with hard tasks, they routinely violated constraints” This—routin

CADENet: Condition-Adaptive Asynchronous Dual-Stream Enhancement Network for Adverse Weather Perception in Autonomous Driving

SafetyDGX agent

arXiv:2605.19837v1 Announce Type: cross Abstract: Adverse weather (rain, fog, sand, and snow) degrades camera-based object detection in autonomous vehicles. Existing enhancement-then-detect approaches

CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

SafetyDGX agent

arXiv:2605.19436v1 Announce Type: cross Abstract: When a model produces a correct solution under reinforcement learning with verifiable rewards (RLVR), every token receives the same reward signal rega

Compliant Explicit Reference Governor for Contact Friendly Robotic Manipulators

SafetyDGX agent

arXiv:2504.09188v2 Announce Type: replace Abstract: This paper introduces the Compliant Explicit Reference Governor (CERG), a modular reference management system that enables robots to interact physic

Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations

Model ReleasesDGX agent

arXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

SafetyDGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support

Model ReleasesDGX agent

arXiv:2605.19826v1 Announce Type: new Abstract: Operators of safety-critical industrial processes increasingly rely on digital twins to screen control interventions, but such simulators rarely carry c

Graph Neural Planning and Predictive Control for Multi-Robot Communication-Constrained Unlabeled Motion Planning

SafetyDGX agent

arXiv:2605.19209v1 Announce Type: new Abstract: The multi-robot unlabeled motion planning problem of concurrently assigning robots to goals and generating safe trajectories is central in many collabor

Hamilton--Jacobi Reachability for Spacecraft Collision Avoidance

SafetyDGX agent

arXiv:2605.20138v1 Announce Type: new Abstract: This article presents a Hamilton--Jacobi (HJ) reachability framework for a two--satellite collision avoidance problem operating in the same circular orb

KIO-planner: Attention-Guided Single-Stage Motion Planning with Dual Mapping for UAV Navigation

Local AiDGX agent

arXiv:2605.19703v1 Announce Type: new Abstract: Autonomous UAV flight in confined, wall-dense environments requires low-latency and reliable motion planning under strict safety constraints. Traditiona

Knowing When Not to Predict: Self Supervised Learning and Abstention for Safer DR Screening

SafetyDGX agent

arXiv:2605.19133v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is now a standard way to pretrain medical image models, but performance is still mostly judged by downstream accuracy.

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

SafetyDGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

Probabilistic Recursively Feasible Motion Planning Under Uncertain Environments

SafetyDGX agent

arXiv:2605.19015v1 Announce Type: cross Abstract: Safe motion planning in uncertain, time-varying environments is challenging because the safe region can change unpredictably across planning steps, of

Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models

SafetyDGX agent

arXiv:2605.19663v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are becoming the cornerstone of high-level reasoning for robotic automation, enabling robots to parse natural language com

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

SafetyDGX agent

arXiv:2605.19940v1 Announce Type: new Abstract: Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cum

19 May 2026

Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?

SafetyDGX agent

arXiv:2605.16354v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as automated evaluators of AI systems, including in high-stakes applications. In this role, LLMs ar

Code as Agent Harness

SafetyDGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems

SafetyDGX agent

arXiv:2605.17909v1 Announce Type: new Abstract: As autonomous agentic systems scale across regulated critical infrastructures, the lack of mechanistic, hardware-rooted enforcement for high-frequency p

Event-Grounded Sparse Autoencoders for Vision-Language-Action Policies

SafetyDGX agent

arXiv:2605.17204v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies translate language and visual inputs into robot actions, where their hidden representations directly shape close

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

SafetyDGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

For a long time, academic researchers being at the cutting edge of new technologies has been a great social equilibrium. Neutral, unbiased t…

SafetyDGX agent

For a long time, academic researchers being at the cutting edge of new technologies has been a great social equilibrium. Neutral, unbiased technologists have been the people to spread new ideas to the

Guided Reinforcement Learning for Omnidirectional 3D Jumping in Quadruped Robots

SafetyDGX agent

arXiv:2507.16481v3 Announce Type: replace Abstract: Jumping poses a significant challenge for quadruped robots, despite being crucial for many operational scenarios. While optimisation methods exist f

Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes

SafetyDGX agent

arXiv:2605.16268v1 Announce Type: cross Abstract: Banks receive millions of reports of fraud, scams, and disputed transactions every year, making it challenging to accurately direct customers to the a

Learning-Based Adaptive Control for Surgical Robotic Exposure Task on Deformable Tissues

SafetyDGX agent

arXiv:2605.17927v1 Announce Type: new Abstract: In various surgical procedures, regions of interest (ROIs) such as organs or lesions are often occluded by overlying tissues, requiring surgeons to achi

On Safer Reinforcement Learning for Sedation and Analgesia in Intensive Care

SafetyDGX agent

arXiv:2601.23154v2 Announce Type: replace-cross Abstract: Pain management in intensive care usually involves complex trade-offs, since both inadequate and excessive treatment can compromise patient sa

Online Learnability of Chain-of-Thought Verifiers: Soundness and Completeness Trade-offs

SafetyDGX agent

arXiv:2603.03538v3 Announce Type: replace Abstract: Large Language Models (LLMs) with chain-of-thought generation have demonstrated great potential for solving complex reasoning and planning tasks. Ho

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures

SafetyDGX agent

arXiv:2605.16551v1 Announce Type: new Abstract: Evaluating LLM-based agents remains challenging because identifying meaningful failure cases often requires substantial human effort to design realistic

Real2Sim via Active Perception with Behavior Trees Automatically Generated by VLMs

SafetyDGX agent

arXiv:2601.08454v2 Announce Type: replace Abstract: Constructing physically accurate simulation environments (Real2Sim) traditionally relies on manual system identification or rigid, exhaustive explor

SuReNav: Superpixel Graph-based Constraint Relaxation for Navigation in Over-constrained Environments

SafetyDGX agent

arXiv:2602.06807v2 Announce Type: replace-cross Abstract: We address the over-constrained planning problem in semi-static environments. The planning objective is to find a best-effort solution that av

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents

SafetyDGX agent

arXiv:2605.17320v1 Announce Type: cross Abstract: Computer-use agents increasingly operate inside live personal workspaces, where their actions can modify files, applications, GUI state, credentials,

Temporal Task Diversity: Inductive Biases Under Non-Stationarity in Synthetic Sequence Modelling

SafetyDGX agent

arXiv:2605.18281v1 Announce Type: new Abstract: Modern deep learning science often assumes that neural networks learn from a fixed data distribution. However, many practically important learning probl

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

SafetyDGX agent

arXiv:2605.17480v1 Announce Type: new Abstract: Multi-agent systems extend large language models (LLMs) by decomposing tasks among specialized agents, but their distributed decision process creates ne

← Previous
1…4344454647…240
Next →