AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Towards Healthy Evolution: Exploring the Role and Mechanisms of Human-Agent Interaction in Self-Evolving Systems

DGX agent

arXiv:2606.06114v1 Announce Type: new Abstract: Self-evolving agents improve through continual self-play and self-generated learning signals, but autonomous evolution can also cause capability degrada

safetyarxiv-cs-ai
6 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Real-Time Threat Detection from Surveillance Cameras using Machine Learning

DGX agent

arXiv:2606.05708v1 Announce Type: new Abstract: Ensuring public safety in densely populated urban environments remains a critical challenge, necessitating the deployment of intelligent and automated v

safetyarxiv-cs-cv
5 Jun 2026
Safety

Expert-Aware Refusal Steering

DGX agent

arXiv:2606.04160v1 Announce Type: new Abstract: Safety alignment in instruction-tuned large language models (LLMs) depends on a model's ability to reliably refuse to respond to harmful or disallowed r

safetyarxiv-cs-cl
4 Jun 2026
Safety

Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms

DGX agent

arXiv:2606.04767v1 Announce Type: cross Abstract: The robustness of deep neural networks is crucial for safety-critical deployments, yet existing evaluation methods are often attack-dependent and lack

safetyarxiv-cs-cv
4 Jun 2026
Safety

Testing Neural Networks via Bayesian-Guided Exploration of Decision Landscapes

DGX agent

arXiv:2606.04314v1 Announce Type: new Abstract: As neural networks are increasingly deployed in safety-critical domains, testing is essential to evaluate and improve their reliability. Existing testin

safetyarxiv-cs-lg
4 Jun 2026
Safety

Assessing Region-Level EEG Contributions to Cognitive Workload Prediction

DGX agent

arXiv:2606.02598v1 Announce Type: new Abstract: Accurate and generalizable estimation of cognitive workload from electroencephalography (EEG) is critical for human-centered and safety-critical systems

safetyarxiv-cs-lg
3 Jun 2026
Safety

D-Judge: Disrupting Multi-Turn Jailbreaks using Semantics-Preserving Output Rewriting

DGX agent

arXiv:2606.02640v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to large language model (LLM) safety because they exploit feedback from auxiliary judge models to i

safetyarxiv-cs-ai
3 Jun 2026
Safety

Glass Box at Orbit: A Constitutional AI Verification Framework for Trustworthy Autonomous CubeSat Intelligence

DGX agent

arXiv:2606.02967v1 Announce Type: cross Abstract: The space industry is quietly building toward something nobody has fully reckoned with: orbital data centers running thousands of autonomous AI worklo

safetyarxiv-cs-ai
3 Jun 2026
Safety

LAP: An Agent-to-Instrument Protocol for Autonomous Science

DGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Learning Power Flow with Confidence: A Probabilistic Guarantee Framework for Voltage Risk

DGX agent

arXiv:2308.07867v4 Announce Type: replace-cross Abstract: The absence of formal performance guarantees in machine learning (ML) has limited its adoption for safety-critical power system applications,

safetyarxiv-cs-lg
3 Jun 2026
Safety

SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural Information Theory

DGX agent

arXiv:2511.16275v4 Announce Type: replace-cross Abstract: Reliable uncertainty quantification (UQ) is essential for deploying large language models (LLMs) in safety-critical scenarios, as it enables t

safetyarxiv-cs-ai
3 Jun 2026
Safety

Adversarial Attacks on Robot Localization Systems via Deep Feature Perturbation

DGX agent

arXiv:2606.01892v1 Announce Type: new Abstract: Robot localization systems are critical for autonomous navigation and safety. Adversarial perturbations can mislead these systems, resulting in mislocal

safetyarxiv-cs-cv
2 Jun 2026
Safety

CEAR: Certified Ensemble Adversarial Robustness in DNNs

DGX agent

arXiv:2606.01437v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) are highly susceptible to adversarial perturbations, leading to extensive research on robustness for safety-critical appli

safetyarxiv-cs-ai
2 Jun 2026
Safety

Constrained Whole-Body Tracking for Humanoid Robots

DGX agent

arXiv:2606.00374v1 Announce Type: new Abstract: Recent advances in reinforcement learning (RL) have demonstrated impressive whole-body agility for humanoid robots, yet ensuring safety and satisfying c

safetyarxiv-cs-ro
2 Jun 2026
Safety

DriveAnchor: Progressive Anchor-based Flow Learning for Autonomous Driving Planning

DGX agent

arXiv:2606.00519v1 Announce Type: new Abstract: We present DriveAnchor, a three-stage framework for autonomous driving planning that achieves behavioral diversity, controllability, and safety in a com

safetyarxiv-cs-ro
2 Jun 2026
Safety

From Cues to Horizons: Dynamic Risk Horizon Profiling for Trajectory Prediction

DGX agent

arXiv:2606.00857v1 Announce Type: cross Abstract: Accurate and reliable vehicle trajectory prediction is essential for safe autonomous driving. Recent studies have incorporated safety risk into trajec

safetyarxiv-cs-ai
2 Jun 2026
Safety

IstGPT: LLM-based Anomaly Detection for Spatial-Temporal Graph in Industrial Systems

DGX agent

arXiv:2606.01691v1 Announce Type: cross Abstract: Industrial Internet systems face increasing threats from sophisticated industrial control system (ICS) attacks, resulting in critical safety incidents

safetyarxiv-cs-lg
2 Jun 2026
Safety

LFA: Layer Feature Attention for Run-Time Introspection of 2D Object Detectors in Automated Driving

DGX agent

arXiv:2606.00372v1 Announce Type: new Abstract: Reliable object detection is critical for automated driving, yet even state-of-the-art detectors inevitably make errors that can compromise safety. Intr

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

Plan-R1: Safe and Feasible Trajectory Planning as Language Modeling

DGX agent

arXiv:2505.17659v4 Announce Type: replace-cross Abstract: Safe and feasible trajectory planning is critical for real-world autonomous driving systems. However, existing learning-based planners rely he

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-Action Models

DGX agent

arXiv:2601.14323v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed in safety-critical robotic applications, yet their security vulnerabilities rema

safetyarxiv-cs-ai
2 Jun 2026
Safety

Train, Test, Re-evaluate: Schedule-Sensitive Evaluation of Generative Data for Hand Detection

DGX agent

arXiv:2606.01896v1 Announce Type: cross Abstract: Generated (or synthetic) image data is increasingly used to augment or replace real training datasets when target imagery is scarce, expensive, or bia

safetyarxiv-cs-ai
2 Jun 2026
Safety

Simulation of collision avoidance behavior in crowd movement by data-driven approach

DGX agent

arXiv:2605.31210v1 Announce Type: cross Abstract: Crowd movement simulation is essential for pedestrian safety management and facility layout optimization. Data-driven models enhance trajectory predic

safetyarxiv-cs-ai
1 Jun 2026
Safety

Unsupervised Defect Detection for Surgical Instruments

DGX agent

arXiv:2509.21561v2 Announce Type: replace Abstract: Ensuring the safety of surgical instruments requires reliable detection of visual defects. However, manual inspection is prone to error, and existin

safetyarxiv-cs-cv
1 Jun 2026
Safety

DefSynUS: Real-time Patient-specific Intrahepatic Vessel Identification via Deformation-Aware CT-US Domain Adaptation

DGX agent

arXiv:2605.29570v1 Announce Type: new Abstract: Purpose: Laparoscopic ultrasound (LUS) enhances the safety of liver surgery by visualizing intrahepatic vessels in real-time. Still, vessel identificati

safetyarxiv-cs-cv
29 May 2026
Safety

Low-Magnification SEM May Suffice: Interpretable Deep Learning for Multi-Scale Fracture-Cause Classification in Zirconia-Toughened Alumina

DGX agent

arXiv:2605.29798v1 Announce Type: new Abstract: Reliable identification of fracture origins in alumina matrix composite hip and knee implants is critical for quality assurance and patient safety, yet

safetyarxiv-cs-cv
29 May 2026
Safety

Masked Diffusion Modeling for Anomaly Detection

DGX agent

arXiv:2605.30046v1 Announce Type: cross Abstract: Anomaly detection aims to identify samples that deviate from the nominal data distribution and is central to many safety-critical applications. Howeve

safetyarxiv-cs-ai
29 May 2026
Safety

Multi-Resolution End-to-End Deep Neural Network for Optimizing Latency-Accuracy Tradeoff in Autonomous Driving

DGX agent

arXiv:2605.29138v1 Announce Type: cross Abstract: Latency-accuracy tradeoffs are fundamental in real-time applications of deep neural networks (DNNs) for cyber-physical systems. In autonomous driving,

safetyarxiv-cs-ai
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Safety

SigmaMedStat: Temporal Signal Modeling for ICU False Alarm Reduction

DGX agent

arXiv:2605.29236v1 Announce Type: new Abstract: Alarm fatigue in intensive care units (ICUs) is a well documented patient safety crisis. Clinical monitors generate 350 or more alarms per patient per d

safetyarxiv-cs-lg
29 May 2026
Safety

V2XCrafter: Learning to Generate Driving Scene Across Agents

DGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

safetyarxiv-cs-cv
29 May 2026
Safety

AOE: Exhaustive Out-of-Distribution Detection via Recalibrating Outlier Labels

DGX agent

arXiv:2605.28021v1 Announce Type: new Abstract: Out-of-distribution (OOD) detection is essential for deploying machine learning models in open-world and safety-critical scenarios, where test inputs ma

safetyarxiv-cs-lg
28 May 2026
Safety

COTTA: Context-Aware Transfer Adaptation for Trajectory Prediction in Autonomous Driving

DGX agent

arXiv:2604.00402v2 Announce Type: replace-cross Abstract: Developing robust models to accurately predict the trajectories of surrounding agents is fundamental to autonomous driving safety. However, mo

safetyarxiv-cs-ai
28 May 2026
Safety

Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security

DGX agent

arXiv:2605.27823v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to adversarial prompts that exploit semantic ambiguities to bypass safety mechanisms, resulti

safetyarxiv-cs-ai
28 May 2026
Safety

Provably Guaranteed Polytopic Uncertainty Quantification for SLAM

DGX agent

arXiv:2605.28172v1 Announce Type: new Abstract: In safety-critical robotics applications, guaranteed and practical uncertainty quantification (UQ) in perception is vital. Many existing works either of

safetyarxiv-cs-ro
28 May 2026
Safety

Voluntary Collusion with Secret Tools in Competing LLM Agents

DGX agent

arXiv:2605.27593v1 Announce Type: new Abstract: Even when a tool is explicitly described as unfair and harmful to others, ostensibly safety-aligned LLM agents still voluntarily engage in secret collus

safetyarxiv-cs-ai
28 May 2026
Safety

CmIVTP: Cross-modal Interaction-based Vessel Trajectory Prediction for Maritime Intelligence

DGX agent

arXiv:2605.26524v1 Announce Type: cross Abstract: Maritime intelligent transportation systems (MITS) are essential for ensuring navigation safety and efficiency in busy waterways. However, accurate ve

safetyarxiv-cs-ai
27 May 2026
Safety

Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces

DGX agent

arXiv:2605.25352v1 Announce Type: cross Abstract: Deep learning models are vulnerable to adversarial perturbations, raising important concerns for safety-critical deployment. Empirical defenses can ac

safetyarxiv-cs-ai
26 May 2026
Safety

Measuring the Depth of LLM Unlearning via Activation Patching

DGX agent

arXiv:2605.24614v1 Announce Type: cross Abstract: Large language model (LLM) unlearning has emerged as a crucial post-hoc mechanism for privacy protection and AI safety, yet auditing whether target kn

safetyarxiv-cs-ai
26 May 2026
Model Releases

Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection

DGX agent

arXiv:2605.24834v1 Announce Type: cross Abstract: Large language model (LLM) safety classifiers such as Llama Guard are effective at detecting overtly harmful prompts but remain vulnerable to adversar

model-releasesarxiv-cs-ai
26 May 2026
Safety

Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation Simulation

DGX agent

arXiv:2605.24535v1 Announce Type: cross Abstract: Jailbreak prompts can trigger harmful completions on aligned LLMs, In accordance, safety steering has been proposed: test-time activation intervention

safetyarxiv-cs-lg
26 May 2026
Safety

TorchLean: Formalizing Neural Networks in Lean

DGX agent

arXiv:2602.22631v2 Announce Type: replace-cross Abstract: Neural networks are increasingly deployed in scientific, safety critical, and mission critical pipelines, yet verification and analysis are of

safetyarxiv-cs-lg
26 May 2026
Safety

Graph-based Complexity Forecasts in UK En Route Airspace Using Relevant Aircraft Interactions

DGX agent

arXiv:2605.23696v1 Announce Type: new Abstract: Effectively managing Air Traffic Control Officer (ATCO) workload is crucial in maintaining operational safety. Group supervisors use tools that estimate

safetyarxiv-cs-lg
25 May 2026
Safety

Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preference Learning

DGX agent

arXiv:2605.23320v1 Announce Type: new Abstract: Ventilator decision support requires sequential decisions that track evolving physiology and disease trajectories while respecting safety boundaries and

safetyarxiv-cs-ai
25 May 2026
Safety

Relevant Walk Search for Explaining Graph Neural Networks

DGX agent

arXiv:2605.23673v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have become important machine learning tools for graph analysis, and its explainability is crucial for safety, fairness, an

safetyarxiv-cs-lg
25 May 2026
Safety

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians

DGX agent

arXiv:2605.23098v1 Announce Type: new Abstract: Reliable uncertainty estimation is critical for deploying monocular depth deep neural networks (DNNs) in safety-critical robotic systems. Conventional u

safetyarxiv-cs-ro
25 May 2026
Safety

Expectation Consistency Loss: Rethink Confidence Calibration under Covariate Shift

DGX agent

arXiv:2605.21552v1 Announce Type: new Abstract: Confidence calibration for classification models is vital in safety-critical decision-making scenarios and has received extensive attention. General con

safetyarxiv-cs-lg
23 May 2026
Safety

MambaGaze: Bidirectional Mamba with Explicit Missing Data Modeling for Cognitive Load Assessment from Eye-Gaze Tracking Data

DGX agent

arXiv:2605.22775v1 Announce Type: new Abstract: Real-time cognitive load assessment from eye-tracking signals could potentially enable adaptive human-centered-AI such as safety-critical applications s

safetyarxiv-cs-lg
23 May 2026
Safety

The Matching Principle: A Geometric Theory of Loss Functions for Nuisance-Robust Representation Learning

DGX agent

arXiv:2605.22800v1 Announce Type: new Abstract: Robustness, domain adaptation, photometric and occlusion invariance, compositional generalisation, temporal robustness, alignment safety, and classical

safetyarxiv-cs-lg
23 May 2026
← Previous
1…2425262728…257
Next →