AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
7 Jul 2026

Probabilistic Robustness in Medical Image Classification

SafetyDGX agent

arXiv:2607.03797v1 Announce Type: new Abstract: Deep learning (DL) has shown strong performance in medical image classification, but its trustworthy deployment remains challenging in safety-critical c

RL-Ballast: Ship Ballast Water Path Planning and Clog Prediction via Reinforcement Learning

SafetyDGX agent

arXiv:2607.04906v1 Announce Type: new Abstract: Under the Shipping 4.0 paradigm, autonomous and reduced-crew vessels require intelligent internal systems to maintain operational safety and structural

Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control

SafetyDGX agent

arXiv:2603.10938v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning from Human Feedback (RLHF) typically enforces safety through expected cost constraints, but the expectation captur

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Uncertainty Quantification for Regression: A Unified Framework based on kernel scores

SafetyDGX agent

arXiv:2510.25599v2 Announce Type: replace Abstract: Regression tasks, notably in safety-critical domains, require reliable uncertainty quantification, yet the literature remains largely classification

UNDREAM: Bridging Differentiable Rendering and Photorealistic Simulation for End-to-end Adversarial Attacks

SafetyDGX agent

arXiv:2510.16923v3 Announce Type: replace-cross Abstract: Deep learning models deployed in safety critical applications like autonomous driving use simulations to test their robustness against adversa

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

SafetyDGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games

SafetyDGX agent

arXiv:2607.05132v1 Announce Type: cross Abstract: As large language models are deployed as autonomous agents that communicate intentions before acting, a critical safety question is whether agents tha

3 Jul 2026

Rethinking Post-Hoc Calibration in Semantic Segmentation

SafetyDGX agent

arXiv:2607.01902v1 Announce Type: cross Abstract: Reliable confidence estimates are essential in semantic segmentation, especially in safety-critical settings where overconfident errors can mislead do

2 Jul 2026

Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces

SafetyDGX agent

arXiv:2607.00481v1 Announce Type: cross Abstract: Jailbreak attacks remain a critical threat to the safe deployment of large language models (LLMs). While prior work has primarily studied attacks and

Decentralized Geometric Control for Cable-Suspended Payload Transport with Adaptive Mass Estimation

SafetyDGX agent

arXiv:2607.00024v1 Announce Type: new Abstract: Cooperative aerial transport requires controllers that respect nonlinear manifold geometry, operate without centralized coordination, and respect operat

Evaluating Implicit Biases in LLM Reasoning through Logic Grid Puzzles

SafetyDGX agent

arXiv:2511.06160v2 Announce Type: replace Abstract: While recent safety guardrails effectively suppress overtly biased outputs, subtler forms of social bias emerge during complex logical reasoning tas

From Prediction Uncertainty to Conformalized Distance Fields for Safe Motion Planning

SafetyDGX agent

arXiv:2607.00776v1 Announce Type: new Abstract: Safe motion planning in dynamic environments requires reasoning about the uncertainty in predicted obstacle motion without sacrificing real-time perform

From Silos to Systems: Process-Oriented Hazard Analysis for AI Systems

SafetyDGX agent

arXiv:2410.22526v2 Announce Type: replace Abstract: To effectively address potential harms from Artificial Intelligence (AI) systems, it is essential to identify and mitigate system-level hazards. Cur

Learning Expert Strategy for Autonomous Robotic Endovascular Intervention via Decoupled Procedural Execution

SafetyDGX agent

arXiv:2607.00066v1 Announce Type: new Abstract: Endovascular interventions are high-stakes procedures requiring precise device operation within complex and tortuous vascular anatomies. Autonomous endo

Sam Altman calls for US-led international forum to set global AI standards

SafetyDGX agent

Sam Altman is calling for a US-led international forum to set global safety standards for artificial intelligence, arguing that no single country should be left to dominate the technology. In an op-ed

1 Jul 2026

DDIAgents: Mechanism-Conditioned Context Flow for Drug-Drug Interaction Prediction

SafetyDGX agent

arXiv:2606.31085v1 Announce Type: new Abstract: Drug-drug interaction (DDI) prediction is essential for medication safety, yet it requires reasoning over heterogeneous biomedical evidence whose releva

FLARE-AI: Flaw Reporting for AI

SafetyDGX agent

arXiv:2606.31567v1 Announce Type: cross Abstract: Flaw reporting for deployed AI systems is fundamental to identifying system failures and improving AI safety. Yet the AI reporting ecosystem is fragme

Great new episode with @So8res on @PeterMcCormack's podcast, give it a watch! https://www.youtube.com/watch?v=1laTnwAdaLA

SafetyDGX agent

Connor Leahy promoted a podcast episode featuring So8res (a researcher/AI safety figure) on Peter McCormack's podcast, sharing a YouTube link to the full video on X/Twitter. The post encourages viewer

LabGuard: Grounding Natural-Language Laboratory Rules into Runtime Guards for Embodied Laboratory Agents

SafetyDGX agent

arXiv:2606.31045v1 Announce Type: new Abstract: Scientific embodied agents are increasingly capable of carrying out laboratory procedures, but executing these procedures safely in dynamic laboratory e

Stealthy Multi-Task Adversarial Attacks

SafetyDGX agent

arXiv:2411.17936v2 Announce Type: replace-cross Abstract: Deep neural networks are highly vulnerable to adversarial perturbations, raising serious safety concerns in the real-world systems. While prio

30 Jun 2026

ActiveVital: Geometry-Aware Embodied Vital Signs Monitoring for Home Healthcare Robots

SafetyDGX agent

arXiv:2606.30275v1 Announce Type: new Abstract: Home robots require reliable vital signs monitoring to support long-term companionship and safety in daily environments, yet obtaining respiration and h

LaGen: Towards Autoregressive LiDAR Scene Generation

SafetyDGX agent

arXiv:2511.21256v2 Announce Type: replace Abstract: Generative world models for autonomous driving (AD) are of great value in applications such as data augmentation, closed-loop simulation, and safety

VISTA-DZ: Visual Semantic Trajectory Adaptation for Personalized Dilemma Zone Prediction

SafetyDGX agent

arXiv:2606.29548v1 Announce Type: cross Abstract: Driver decision making in the dilemma zone at signalized intersections is safety critical, as vehicles approaching a yellow signal must decide whether

Zero-Label Driving Scenario Complexity Detection via Joint Embedding Predictive Architecture

SafetyDGX agent

arXiv:2606.28383v1 Announce Type: new Abstract: Identifying complex and safety-critical driving scenarios in large unlabelled datasets is an important but expensive problem. Existing approaches rely o

29 Jun 2026

OperatorSHAP: Fast and Accurate Shapley Value Estimation for Neural Operators

SafetyDGX agent

arXiv:2606.28065v1 Announce Type: cross Abstract: Understanding model predictions is essential for physical applications, where outputs often inform safety-critical decisions, such as structural load

PPO-EAL: Exact Augmented Lagrangian Proximal Policy Optimization for Safe Robotic Control

SafetyDGX agent

arXiv:2606.27861v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising solution to accomplish complex robotic control tasks; however, most of the current work ignores t

Reasoning-Enhanced Rare-Event Prediction with Balanced Outcome Correction

SafetyDGX agent

arXiv:2601.16406v2 Announce Type: replace-cross Abstract: Rare-event prediction is critical in domains such as healthcare, finance, reliability engineering, customer support, aviation safety, where po

26 Jun 2026

Chai: Agentic Discovery of Cryptographic Misuse Vulnerabilities

SafetyDGX agent

arXiv:2606.26933v1 Announce Type: cross Abstract: AI-assisted vulnerability discovery has proven effective for bug classes like memory safety, where instrumentation confirms memory violations and effi

HauntAttack: When Attack Follows Reasoning as a Shadow

SafetyDGX agent

arXiv:2506.07031v5 Announce Type: replace-cross Abstract: Emerging Large Reasoning Models (LRMs) consistently excel in mathematical and reasoning tasks, showcasing remarkable capabilities. However, th

RecallRisk-BERT: A Multi-Task Framework for Post-Report Medical Device Recall Triage

SafetyDGX agent

arXiv:2606.27174v1 Announce Type: new Abstract: Medical device recalls are a critical regulatory mechanism for protecting patient safety. The growing volume of FDA recall records presents challenges i

25 Jun 2026

Falcon: Functional Assembly and Language for Compositional Reasoning in X-ray

Model ReleasesDGX agent

arXiv:2606.25701v1 Announce Type: new Abstract: Conventional vision-language models are largely object-centric, focusing on detecting and describing individual entities. In safety-critical X-ray bagga

The Tatoxa System for Text Detoxification in Low-Resource Languages: The Case of Tatar

SafetyDGX agent

arXiv:2606.26015v1 Announce Type: new Abstract: Text detoxification, the automated detection and mitigation of abusive and harmful content, is essential for ensuring the safety of online communities a

Towards Understanding The Calibration Benefits of Sharpness-Aware Minimization

SafetyDGX agent

arXiv:2505.23866v2 Announce Type: replace Abstract: Deep neural networks have been increasingly used in safety-critical applications such as medical diagnosis and autonomous driving. However, many stu

24 Jun 2026

DDStereo: Efficient Dual Decoder Transformers for Stereo 3D Road Anomaly Detection

SafetyDGX agent

arXiv:2606.24805v1 Announce Type: new Abstract: Stereo-based 3D object detection still faces two critical safety challenges: real-time performance and open-set generalization. Existing stereo 3D metho

One Year Later...The Harms Persist, But So Do We!

SafetyDGX agent

arXiv:2606.23884v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) are increasingly used for mental health-related conversations, yet safety safeguards remain inadequate an

23 Jun 2026

An LLM-Explainable DRL Framework for Passenger-Directed Autonomous Driving

SafetyDGX agent

arXiv:2606.20640v1 Announce Type: cross Abstract: Autonomous vehicles offer the potential for safer and more efficient mobility, yet public trust remains limited due to the lack of transparency in the

BayesFP: Posterior Estimation for Flow-Based Policies via Feynman-Kac Sampling

SafetyDGX agent

arXiv:2606.21014v1 Announce Type: new Abstract: Robots must generate trajectories that remain faithful to learned expert behavior while satisfying safety constraints and task-specific objectives speci

Conflict-Aware Switching for CBF-CLF-Based Multi-Goal Navigation

SafetyDGX agent

arXiv:2606.21577v1 Announce Type: new Abstract: Quadratic programs (QPs) using Control Barrier Functions (CBFs) and Control Lyapunov Functions (CLFs) are widely used for safe control in reach-and-avoi

CRAX: Fast Safe Reinforcement Learning Benchmarking

Model ReleasesDGX agent

arXiv:2606.20376v2 Announce Type: replace Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While bench

DBT-Bleed: Dual-Branch Temporal Modeling with Key-Frame Selection for Surgical Bleeding Detection

SafetyDGX agent

arXiv:2606.22829v1 Announce Type: new Abstract: Intraoperative Adverse Events (IAEs) detection is critical for improving surgical safety, with bleeding being among the most frequent events across many

From Driving Videos to Simulatable Scenarios

SafetyDGX agent

arXiv:2606.21993v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) face driving scenarios ranging from routine traffic to rare events. To assess safety it is crucial to reproduce these scenar

My biggest NY-12 competitors aren't on the ballot today: Trump's AI oligarchs. If they can defeat me—the author of the nation's strongest AI…

SafetyDGX agent

My biggest NY-12 competitors aren't on the ballot today: Trump's AI oligarchs. If they can defeat me—the author of the nation's strongest AI safety bill in one of the nation's bluest districts—they ca

NegAS: Negative Label Guided Attention and Scoring for Out-of-Distribution Object Detection with Vision-Language Models

SafetyDGX agent

arXiv:2606.22537v1 Announce Type: new Abstract: Out-of-Distribution (OOD) detection is essential for ensuring the robustness and reliability of object detection systems deployed in safety-critical app

Neural Architecture Distributions: A New Paradigm for Stochastic Segmentation

SafetyDGX agent

arXiv:2606.21061v1 Announce Type: new Abstract: Stochastic segmentation seeks to represent multiple plausible masks for a single image, which is essential in safety- and quality-critical applications

Predicting cognitive load in immersive driving scenarios with a hybrid CNN-RNN model

SafetyDGX agent

arXiv:2408.06350v2 Announce Type: replace-cross Abstract: One debatable issue in traffic safety research is that cognitive load from sec-ondary tasks reduces primary task performance, such as driving.

Safe Few-Step Generation via Velocity Editing

SafetyDGX agent

arXiv:2606.23267v1 Announce Type: new Abstract: Flow matching has recently emerged as a strong paradigm for state-of-the-art text-to-image (T2I) generation, enabling high-quality generation with a sma

SPiralRoll: A Novel Adjustable-Stiffness Underactuated 3-DoF Joint with Torsion Springs for Rolling Robots

SafetyDGX agent

arXiv:2606.22443v1 Announce Type: new Abstract: Compliant mechanisms are important in robotics because they can improve adaptability, safety, and energy efficiency while reducing hardware complexity.

11 Jun 2026

AutoMine Solution for AV2 2026 Scenario Mining Challenge

SafetyDGX agent

arXiv:2606.11874v1 Announce Type: new Abstract: With the development of autonomous driving systems, mining high-value, safety-critical, and planning-relevant scenarios from large-scale driving logs ha

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

SafetyDGX agent

arXiv:2606.11648v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean

Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers

SafetyDGX agent

arXiv:2606.12025v1 Announce Type: new Abstract: Finite element (FE) modeling of safety-critical infrastructure such as bridge barriers requires high-fidelity nonlinear dynamic analysis, yet the curren

JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization

SafetyDGX agent

arXiv:2606.11425v1 Announce Type: cross Abstract: Jailbreak attacks expose persistent safety weaknesses in large language models (LLMs), but existing stateless single-turn methods face a trade-off: ha

One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection

SafetyDGX agent

arXiv:2606.11202v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in domina

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models

SafetyDGX agent

arXiv:2606.11266v1 Announce Type: new Abstract: The cost signal that constrained-RL algorithms optimize against is almost always reactive: the simulator emits a non-zero cost only after a collision ha

10 Jun 2026

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

SafetyDGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

Anthropic and OpenAI did not call for a pause. Read the wording: 'good for the world to have the option', 'possible' to slow down 'when need…

SafetyDGX agent

Anthropic and OpenAI did not call for a pause. Read the wording: 'good for the world to have the option', 'possible' to slow down 'when needed'. This is how they signal safety to one audience, acceler

Anthropic, if they really believe what they say, should show some leadership:

SafetyDGX agent

Anthropic, if they really believe what they say, should show some leadership: 🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even b

Automatic Labelling for Low-Light Pedestrian Detection

SafetyDGX agent

arXiv:2507.02513v4 Announce Type: replace Abstract: Pedestrian detection in RGB images is a key task in pedestrian safety, as the most common sensor in autonomous vehicles and advanced driver assistan

On the Controllability-Fidelity Frontier in Diffusion Editing

SafetyDGX agent

arXiv:2606.09901v1 Announce Type: cross Abstract: Diffusion-based generative models enable powerful image editing capabilities, but achieving precise control while maintaining fidelity and safety rema

9 Jun 2026

A VideoMAE-v2 Approach to Zero-Shot Traffic Accident Anticipation

SafetyDGX agent

arXiv:2606.09542v1 Announce Type: new Abstract: Traffic accident anticipation -- predicting the likelihood of an imminent collision at every frame of a dashcam video -- is safety-critical yet difficul

Automated Framework to Evaluate and Harden LLM System Instructions against Encoding Attacks

SafetyDGX agent

arXiv:2604.01039v2 Announce Type: replace-cross Abstract: System Instructions in Large Language Models (LLMs) are commonly used to enforce safety policies, define agent behavior, and protect sensitive

← Previous
1…2021222324…240
Next →