AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
Safety

Assessing Region-Level EEG Contributions to Cognitive Workload Prediction

DGX agent

arXiv:2606.02598v1 Announce Type: new Abstract: Accurate and generalizable estimation of cognitive workload from electroencephalography (EEG) is critical for human-centered and safety-critical systems

safetyarxiv-cs-lg
3 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

D-Judge: Disrupting Multi-Turn Jailbreaks using Semantics-Preserving Output Rewriting

DGX agent

arXiv:2606.02640v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to large language model (LLM) safety because they exploit feedback from auxiliary judge models to i

safetyarxiv-cs-ai
3 Jun 2026
Safety

Glass Box at Orbit: A Constitutional AI Verification Framework for Trustworthy Autonomous CubeSat Intelligence

DGX agent

arXiv:2606.02967v1 Announce Type: cross Abstract: The space industry is quietly building toward something nobody has fully reckoned with: orbital data centers running thousands of autonomous AI worklo

safetyarxiv-cs-ai
3 Jun 2026
Safety

LAP: An Agent-to-Instrument Protocol for Autonomous Science

DGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Learning Power Flow with Confidence: A Probabilistic Guarantee Framework for Voltage Risk

DGX agent

arXiv:2308.07867v4 Announce Type: replace-cross Abstract: The absence of formal performance guarantees in machine learning (ML) has limited its adoption for safety-critical power system applications,

safetyarxiv-cs-lg
3 Jun 2026
Safety

SeSE: Black-Box Uncertainty Quantification for Large Language Models Based on Structural Information Theory

DGX agent

arXiv:2511.16275v4 Announce Type: replace-cross Abstract: Reliable uncertainty quantification (UQ) is essential for deploying large language models (LLMs) in safety-critical scenarios, as it enables t

safetyarxiv-cs-ai
3 Jun 2026
Safety

Adversarial Attacks on Robot Localization Systems via Deep Feature Perturbation

DGX agent

arXiv:2606.01892v1 Announce Type: new Abstract: Robot localization systems are critical for autonomous navigation and safety. Adversarial perturbations can mislead these systems, resulting in mislocal

safetyarxiv-cs-cv
2 Jun 2026
Safety

CEAR: Certified Ensemble Adversarial Robustness in DNNs

DGX agent

arXiv:2606.01437v1 Announce Type: cross Abstract: Deep Neural Networks (DNNs) are highly susceptible to adversarial perturbations, leading to extensive research on robustness for safety-critical appli

safetyarxiv-cs-ai
2 Jun 2026
Safety

Constrained Whole-Body Tracking for Humanoid Robots

DGX agent

arXiv:2606.00374v1 Announce Type: new Abstract: Recent advances in reinforcement learning (RL) have demonstrated impressive whole-body agility for humanoid robots, yet ensuring safety and satisfying c

safetyarxiv-cs-ro
2 Jun 2026
Safety

DriveAnchor: Progressive Anchor-based Flow Learning for Autonomous Driving Planning

DGX agent

arXiv:2606.00519v1 Announce Type: new Abstract: We present DriveAnchor, a three-stage framework for autonomous driving planning that achieves behavioral diversity, controllability, and safety in a com

safetyarxiv-cs-ro
2 Jun 2026
Safety

From Cues to Horizons: Dynamic Risk Horizon Profiling for Trajectory Prediction

DGX agent

arXiv:2606.00857v1 Announce Type: cross Abstract: Accurate and reliable vehicle trajectory prediction is essential for safe autonomous driving. Recent studies have incorporated safety risk into trajec

safetyarxiv-cs-ai
2 Jun 2026
Safety

IstGPT: LLM-based Anomaly Detection for Spatial-Temporal Graph in Industrial Systems

DGX agent

arXiv:2606.01691v1 Announce Type: cross Abstract: Industrial Internet systems face increasing threats from sophisticated industrial control system (ICS) attacks, resulting in critical safety incidents

safetyarxiv-cs-lg
2 Jun 2026
Safety

LFA: Layer Feature Attention for Run-Time Introspection of 2D Object Detectors in Automated Driving

DGX agent

arXiv:2606.00372v1 Announce Type: new Abstract: Reliable object detection is critical for automated driving, yet even state-of-the-art detectors inevitably make errors that can compromise safety. Intr

safetyarxiv-cs-cv
2 Jun 2026
Model Releases

Plan-R1: Safe and Feasible Trajectory Planning as Language Modeling

DGX agent

arXiv:2505.17659v4 Announce Type: replace-cross Abstract: Safe and feasible trajectory planning is critical for real-world autonomous driving systems. However, existing learning-based planners rely he

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-Action Models

DGX agent

arXiv:2601.14323v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed in safety-critical robotic applications, yet their security vulnerabilities rema

safetyarxiv-cs-ai
2 Jun 2026
Safety

Train, Test, Re-evaluate: Schedule-Sensitive Evaluation of Generative Data for Hand Detection

DGX agent

arXiv:2606.01896v1 Announce Type: cross Abstract: Generated (or synthetic) image data is increasingly used to augment or replace real training datasets when target imagery is scarce, expensive, or bia

safetyarxiv-cs-ai
2 Jun 2026
Safety

Simulation of collision avoidance behavior in crowd movement by data-driven approach

DGX agent

arXiv:2605.31210v1 Announce Type: cross Abstract: Crowd movement simulation is essential for pedestrian safety management and facility layout optimization. Data-driven models enhance trajectory predic

safetyarxiv-cs-ai
1 Jun 2026
Safety

Unsupervised Defect Detection for Surgical Instruments

DGX agent

arXiv:2509.21561v2 Announce Type: replace Abstract: Ensuring the safety of surgical instruments requires reliable detection of visual defects. However, manual inspection is prone to error, and existin

safetyarxiv-cs-cv
1 Jun 2026
Safety

DefSynUS: Real-time Patient-specific Intrahepatic Vessel Identification via Deformation-Aware CT-US Domain Adaptation

DGX agent

arXiv:2605.29570v1 Announce Type: new Abstract: Purpose: Laparoscopic ultrasound (LUS) enhances the safety of liver surgery by visualizing intrahepatic vessels in real-time. Still, vessel identificati

safetyarxiv-cs-cv
29 May 2026
Safety

Low-Magnification SEM May Suffice: Interpretable Deep Learning for Multi-Scale Fracture-Cause Classification in Zirconia-Toughened Alumina

DGX agent

arXiv:2605.29798v1 Announce Type: new Abstract: Reliable identification of fracture origins in alumina matrix composite hip and knee implants is critical for quality assurance and patient safety, yet

safetyarxiv-cs-cv
29 May 2026
Safety

Masked Diffusion Modeling for Anomaly Detection

DGX agent

arXiv:2605.30046v1 Announce Type: cross Abstract: Anomaly detection aims to identify samples that deviate from the nominal data distribution and is central to many safety-critical applications. Howeve

safetyarxiv-cs-ai
29 May 2026
Safety

Multi-Resolution End-to-End Deep Neural Network for Optimizing Latency-Accuracy Tradeoff in Autonomous Driving

DGX agent

arXiv:2605.29138v1 Announce Type: cross Abstract: Latency-accuracy tradeoffs are fundamental in real-time applications of deep neural networks (DNNs) for cyber-physical systems. In autonomous driving,

safetyarxiv-cs-ai
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Safety

SigmaMedStat: Temporal Signal Modeling for ICU False Alarm Reduction

DGX agent

arXiv:2605.29236v1 Announce Type: new Abstract: Alarm fatigue in intensive care units (ICUs) is a well documented patient safety crisis. Clinical monitors generate 350 or more alarms per patient per d

safetyarxiv-cs-lg
29 May 2026
Safety

V2XCrafter: Learning to Generate Driving Scene Across Agents

DGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

safetyarxiv-cs-cv
29 May 2026
Safety

AOE: Exhaustive Out-of-Distribution Detection via Recalibrating Outlier Labels

DGX agent

arXiv:2605.28021v1 Announce Type: new Abstract: Out-of-distribution (OOD) detection is essential for deploying machine learning models in open-world and safety-critical scenarios, where test inputs ma

safetyarxiv-cs-lg
28 May 2026
Safety

COTTA: Context-Aware Transfer Adaptation for Trajectory Prediction in Autonomous Driving

DGX agent

arXiv:2604.00402v2 Announce Type: replace-cross Abstract: Developing robust models to accurately predict the trajectories of surrounding agents is fundamental to autonomous driving safety. However, mo

safetyarxiv-cs-ai
28 May 2026
Safety

Disentangling Adversarial Prompts: A Semantic-Graph Defense for Robust LLM Security

DGX agent

arXiv:2605.27823v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly vulnerable to adversarial prompts that exploit semantic ambiguities to bypass safety mechanisms, resulti

safetyarxiv-cs-ai
28 May 2026
Safety

OpenAI’s Frontier Governance Framework

DGX agent

OpenAI's Frontier Governance Framework outlines the organization's approach to managing risks associated with advanced AI systems, including safety, security, and responsible deployment practices. The

safetyopenai
28 May 2026
Safety

Provably Guaranteed Polytopic Uncertainty Quantification for SLAM

DGX agent

arXiv:2605.28172v1 Announce Type: new Abstract: In safety-critical robotics applications, guaranteed and practical uncertainty quantification (UQ) in perception is vital. Many existing works either of

safetyarxiv-cs-ro
28 May 2026
Safety

Voluntary Collusion with Secret Tools in Competing LLM Agents

DGX agent

arXiv:2605.27593v1 Announce Type: new Abstract: Even when a tool is explicitly described as unfair and harmful to others, ostensibly safety-aligned LLM agents still voluntarily engage in secret collus

safetyarxiv-cs-ai
28 May 2026
Safety

CmIVTP: Cross-modal Interaction-based Vessel Trajectory Prediction for Maritime Intelligence

DGX agent

arXiv:2605.26524v1 Announce Type: cross Abstract: Maritime intelligent transportation systems (MITS) are essential for ensuring navigation safety and efficiency in busy waterways. However, accurate ve

safetyarxiv-cs-ai
27 May 2026
Safety

Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces

DGX agent

arXiv:2605.25352v1 Announce Type: cross Abstract: Deep learning models are vulnerable to adversarial perturbations, raising important concerns for safety-critical deployment. Empirical defenses can ac

safetyarxiv-cs-ai
26 May 2026
Safety

Measuring the Depth of LLM Unlearning via Activation Patching

DGX agent

arXiv:2605.24614v1 Announce Type: cross Abstract: Large language model (LLM) unlearning has emerged as a crucial post-hoc mechanism for privacy protection and AI safety, yet auditing whether target kn

safetyarxiv-cs-ai
26 May 2026
Model Releases

Reflect-Guard: Enhancing LLM Safeguards against Adversarial Prompts via Logical Self-Reflection

DGX agent

arXiv:2605.24834v1 Announce Type: cross Abstract: Large language model (LLM) safety classifiers such as Llama Guard are effective at detecting overtly harmful prompts but remain vulnerable to adversar

model-releasesarxiv-cs-ai
26 May 2026
Safety

Steering Beyond the Support: Adversarial Training on Unsupervised Jailbroken Activation Simulation

DGX agent

arXiv:2605.24535v1 Announce Type: cross Abstract: Jailbreak prompts can trigger harmful completions on aligned LLMs, In accordance, safety steering has been proposed: test-time activation intervention

safetyarxiv-cs-lg
26 May 2026
Safety

TorchLean: Formalizing Neural Networks in Lean

DGX agent

arXiv:2602.22631v2 Announce Type: replace-cross Abstract: Neural networks are increasingly deployed in scientific, safety critical, and mission critical pipelines, yet verification and analysis are of

safetyarxiv-cs-lg
26 May 2026
Safety

Graph-based Complexity Forecasts in UK En Route Airspace Using Relevant Aircraft Interactions

DGX agent

arXiv:2605.23696v1 Announce Type: new Abstract: Effectively managing Air Traffic Control Officer (ATCO) workload is crucial in maintaining operational safety. Group supervisors use tools that estimate

safetyarxiv-cs-lg
25 May 2026
Safety

Human-in-the-Loop Multi-Agent Ventilator Decision Support with Contextual Bandit Preference Learning

DGX agent

arXiv:2605.23320v1 Announce Type: new Abstract: Ventilator decision support requires sequential decisions that track evolving physiology and disease trajectories while respecting safety boundaries and

safetyarxiv-cs-ai
25 May 2026
Safety

People are often confused that I am against the framing of 'tool AI' This is hands down the best post explaining (some of) the issues with t…

DGX agent

People are often confused that I am against the framing of 'tool AI' This is hands down the best post explaining (some of) the issues with the term. Give it a read! Many in AI safety advocacy argue th

safetyconnor-leahy--x
25 May 2026
Safety

Relevant Walk Search for Explaining Graph Neural Networks

DGX agent

arXiv:2605.23673v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have become important machine learning tools for graph analysis, and its explainability is crucial for safety, fairness, an

safetyarxiv-cs-lg
25 May 2026
Safety

The Pope rightly warns that AI must serve human dignity, not become a tool of domination or exclusion. But if we hand governments sweeping p…

DGX agent

The Pope rightly warns that AI must serve human dignity, not become a tool of domination or exclusion. But if we hand governments sweeping power over AI development in the name of safety, how do we pr

safetyclem-delangue--x
25 May 2026
Safety

UfM*: Uncertainty from Motion* for DNN Depth Estimation Using Gaussians

DGX agent

arXiv:2605.23098v1 Announce Type: new Abstract: Reliable uncertainty estimation is critical for deploying monocular depth deep neural networks (DNNs) in safety-critical robotic systems. Conventional u

safetyarxiv-cs-ro
25 May 2026
Safety

even @geohotz is starting to sound like me 🤣

DGX agent

Gary Marcus humorously notes that George Hotz, an AI researcher and entrepreneur, is beginning to echo Marcus's own views or criticisms, likely regarding AI safety, limitations, or technical concerns.

safetygary-marcus--x
24 May 2026
Safety

Expectation Consistency Loss: Rethink Confidence Calibration under Covariate Shift

DGX agent

arXiv:2605.21552v1 Announce Type: new Abstract: Confidence calibration for classification models is vital in safety-critical decision-making scenarios and has received extensive attention. General con

safetyarxiv-cs-lg
23 May 2026
Safety

MambaGaze: Bidirectional Mamba with Explicit Missing Data Modeling for Cognitive Load Assessment from Eye-Gaze Tracking Data

DGX agent

arXiv:2605.22775v1 Announce Type: new Abstract: Real-time cognitive load assessment from eye-tracking signals could potentially enable adaptive human-centered-AI such as safety-critical applications s

safetyarxiv-cs-lg
23 May 2026
Safety

The Matching Principle: A Geometric Theory of Loss Functions for Nuisance-Robust Representation Learning

DGX agent

arXiv:2605.22800v1 Announce Type: new Abstract: Robustness, domain adaptation, photometric and occlusion invariance, compositional generalisation, temporal robustness, alignment safety, and classical

safetyarxiv-cs-lg
23 May 2026
Safety

Visibility nowcasting in South Korea: a machine learning approach to class imbalance and distribution shift

DGX agent

arXiv:2605.21507v1 Announce Type: cross Abstract: Atmospheric visibility is a critical variable for transportation safety and air quality management, however, accurate prediction remains challenging d

safetyarxiv-cs-lg
23 May 2026
← Previous
1…2728293031…299
Next →