AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

INTENT: An LSTM Framework for Vehicle Intention Prediction in Intersection Scenarios with Comprehensive Ablation Analysis

DGX agent

arXiv:2607.08316v1 Announce Type: new Abstract: Vehicle intention prediction is a pivotal aspect in the agility and safety of autonomous vehicles in all driving scenarios; if genuine enhancement of au

safetyarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

LipSSD: Lipschitz-Constrained Single-Shot Detection for Adversarially Robust Object Detection

DGX agent

arXiv:2607.06592v1 Announce Type: cross Abstract: Object detectors have many applications in safety-critical systems, but they are known to be sensitive to worst-case perturbations such as adversarial

safetyarxiv-cs-ai
9 Jul 2026
Safety

Explainable Reinforcement Learning for Adaptive Traffic Signal Control

DGX agent

arXiv:2607.03703v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a powerful paradigm for adaptive traffic signal control. However, in safety-critical infrastructure like traf

safetyarxiv-cs-ai
7 Jul 2026
Safety

Hope for the Best, Prepare for the Worst: Occlusion-Aware Contingency Planning for Autonomous Vehicles

DGX agent

arXiv:2607.03155v1 Announce Type: new Abstract: The deployment of autonomous vehicles in urban environments introduces significant safety challenges, particularly in scenarios with occlusions, where c

safetyarxiv-cs-ro
7 Jul 2026
Safety

Resolving Primitive-Sharing Ambiguity in Long-Tailed TLS-Based Industrial MEP Point Cloud Segmentation via Spatial Context Constraints

DGX agent

arXiv:2601.19128v2 Announce Type: replace Abstract: In terrestrial laser scanning (TLS)-based mechanical, electrical, and plumbing (MEP) point cloud segmentation, safety-critical components such as re

safetyarxiv-cs-cv
7 Jul 2026
Safety

Behind the Refusal: Determining Guardrail Activation via Behavioral Monitoring

DGX agent

arXiv:2607.02121v1 Announce Type: cross Abstract: As Large Language Models (LLMs) and agentic systems become integrated into real-world applications, ensuring their safety and security is critical. Gu

safetyarxiv-cs-ai
3 Jul 2026
Safety

Lightweight Safe Reinforcement Learning for End-to-End UAV Navigation

DGX agent

arXiv:2607.01794v1 Announce Type: cross Abstract: With the rapid development of autonomous aerial systems, Unmanned Aerial Vehicles (UAVs) are increasingly deployed in applications such as inspection,

safetyarxiv-cs-ai
3 Jul 2026
Safety

Robust Operational Space Control with Conformal Disturbance Bounds for Safe Redundant Manipulation

DGX agent

arXiv:2607.00424v1 Announce Type: new Abstract: Redundant robotic manipulators operating in constrained and human-interactive environments require accurate task-space tracking together with rigorous s

safetyarxiv-cs-ro
2 Jul 2026
Safety

On Optimizing Multimodal Jailbreaks for Spoken Language Models

DGX agent

arXiv:2603.19127v2 Announce Type: replace Abstract: As Spoken Language Models (SLMs) integrate speech and text modalities, they inherit the safety vulnerabilities of their LLM backbone while introduci

safetyarxiv-cs-lg
1 Jul 2026
Safety

Reward Redistribution for CVaR MDPs using a Bellman Operator on L-infinity

DGX agent

arXiv:2602.03778v2 Announce Type: replace-cross Abstract: Tail-end risk measures such as static conditional value-at-risk (CVaR) are used in safety-critical applications to prevent rare, yet catastrop

safetyarxiv-cs-ai
1 Jul 2026
Safety

A Gravitational Interpretation of Fine-Tuning Reversion

DGX agent

arXiv:2606.28525v1 Announce Type: cross Abstract: Fine-tuning on harmless data can partially undo behaviors acquired earlier in training. Safety can erode under benign post-alignment updates, unlearne

safetyarxiv-cs-ai
30 Jun 2026
Safety

Not Just How Much, But Where: Decomposing Epistemic Uncertainty into Per-Class Contributions

DGX agent

arXiv:2602.21160v4 Announce Type: replace-cross Abstract: In safety-critical classification, the cost of failure is often asymmetric, yet Bayesian deep learning summarises epistemic uncertainty with a

safetyarxiv-cs-lg
30 Jun 2026
Safety

Sparse Autoencoders are Capable LLM Jailbreak Mitigators

DGX agent

arXiv:2602.12418v2 Announce Type: replace-cross Abstract: Jailbreak attacks remain a persistent threat to large language model safety. We propose Context-Conditioned Delta Steering (CC-Delta), an SAE-

safetyarxiv-cs-cl
30 Jun 2026
Safety

The Undecidability of Artificial General Intelligence (AGI) Alignment

DGX agent

arXiv:2606.28639v1 Announce Type: cross Abstract: This article establishes the foundational mathematical limits of Artificial General Intelligence (AGI) safety, proving that the core barrier is not th

safetyarxiv-cs-ai
30 Jun 2026
Safety

Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking

DGX agent

arXiv:2606.26205v1 Announce Type: new Abstract: Patients increasingly seek medication information online, yet safety knowledge for psychiatric drugs is split between regulatory adverse-event records,

safetyarxiv-cs-ai
26 Jun 2026
Safety

A Survey of Toxicity Detection and Mitigation Strategies for Multilingual Language Models

DGX agent

arXiv:2606.25380v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed across languages, but their safety behavior remains uneven across linguistic and cultural context

safetyarxiv-cs-cl
25 Jun 2026
Safety

ARTOO-DARTU: Studying AR-HRC With AR Obstruction Mitigation During a Warehouse Task

DGX agent

arXiv:2606.25202v1 Announce Type: cross Abstract: Human-robot collaboration (HRC) often requires robot intentions and internal states to be conveyed to users for task efficiency and safety. Recently,

safetyarxiv-cs-ro
25 Jun 2026
Safety

Conformal Recovery-Deadline Certificates for Runtime Assurance of Adapting Controllers

DGX agent

arXiv:2606.25371v1 Announce Type: cross Abstract: Runtime assurance (RTA) protects a safety-critical system by switching from an advanced controller to a verified safe controller when a monitored cond

safetyarxiv-cs-ai
25 Jun 2026
Safety

A UAV-Based Multi-Modal Vision System for Automated Sideslope Deformation Monitoring and Hazard Detection

DGX agent

arXiv:2606.20681v1 Announce Type: new Abstract: Slope hazards constitute a major safety threat to expressway infrastructure, and their evolution is typically manifested as slow surface deformation. Co

safetyarxiv-cs-cv
23 Jun 2026
Safety

HumanHalo -- Safe and Efficient 3D Navigation Among Humans via Minimally Conservative MPC

DGX agent

arXiv:2510.17525v3 Announce Type: replace Abstract: Safe and efficient robotic navigation among humans is essential for integrating robots into everyday environments. Most existing approaches focus on

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

Alignment Collapse Under KV Cache Quantization: Diagnosis and Mitigation

DGX agent

arXiv:2606.09864v1 Announce Type: cross Abstract: Key-value (KV) cache quantization is widely used to reduce Large Language Model (LLM) inference memory, yet existing evaluations solely focus on measu

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Can Data Work be Reparative?

DGX agent

arXiv:2606.09408v1 Announce Type: cross Abstract: We present an ethnographic study of an alternative approach to data work, developed by a civic-tech initiative that builds datasets for training and b

safetyarxiv-cs-ai
9 Jun 2026
Safety

Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis

DGX agent

arXiv:2606.09178v1 Announce Type: cross Abstract: Multilingual safety evaluation of large language models (LLMs) has predominantly relied on direct translation (DT) of English benchmarks into target l

safetyarxiv-cs-ai
9 Jun 2026
Safety

Diffuse AI Control on Fuzzy Tasks

DGX agent

arXiv:2606.08892v1 Announce Type: new Abstract: AI models deployed in critical domains, such as AI safety research, may subtly sabotage our efforts due to misalignment. Diffuse AI Control is a subfiel

safetyarxiv-cs-lg
9 Jun 2026
Safety

A Causal Probabilistic Framework for Perception-Informed Closed-Loop Simulation of Autonomous Driving

DGX agent

arXiv:2606.07186v1 Announce Type: new Abstract: Software-in-the-loop (SIL) simulation is a cornerstone for the validation of modern automotive safety functions. However, many current frameworks utiliz

safetyarxiv-cs-ro
8 Jun 2026
Safety

Re-imagining ISO 26262 in the Age of Autonomous Vehicles: Enhancing Controllability through Transferability and Predictability

DGX agent

arXiv:2606.07437v1 Announce Type: cross Abstract: The ISO 26262 standard defines functional safety for road vehicles through risk assessments based on Severity, Exposure, and Controllability, grounded

safetyarxiv-cs-ai
8 Jun 2026
Safety

Unified Safe In-context Image Generation in Multimodal Diffusion Transformers via Restricting Unsafe Information Flows

DGX agent

arXiv:2606.06875v1 Announce Type: new Abstract: Diffusion transformers (DiTs) equipped with multimodal attention (MM-Attn) have become a dominant paradigm for image generation. However, preventing the

safetyarxiv-cs-cv
8 Jun 2026
Safety

From Risk Classification to Action Plan Remediation: A Guardrail Feedback Driven Framework for LLM Agents

DGX agent

arXiv:2606.05805v1 Announce Type: new Abstract: LLM-based guardrails typically safeguard agents by evaluating proposed actions or inputs before execution, producing safety signals such as binary allow

safetyarxiv-cs-ai
6 Jun 2026
Safety

Where Should Knowledge Enter? A Layered Framework for Knowledge Infusion in Multimodal Iterative Generative Mo

DGX agent

arXiv:2606.06356v1 Announce Type: new Abstract: Multimodal generative models produce fluent outputs but remain unreliable when generation must respect structured, domain-specific, or safety-critical k

safetyarxiv-cs-ai
6 Jun 2026
Safety

Drishti AI-Event Guardian: An Intelligent Real-Time Crowd Monitoring and Emergency Response System for Mass Gathering Events

DGX agent

arXiv:2606.05185v1 Announce Type: cross Abstract: Mass gathering events are associated with critical safety incidents caused by insufficient crowd monitoring and inadequate emergency response coordina

safetyarxiv-cs-cv
5 Jun 2026
Safety

Instance-Level Post Hoc Uncertainty Quantification in Object Detection

DGX agent

arXiv:2606.04656v1 Announce Type: cross Abstract: Object detection is a safety-critical component of autonomous driving. It is essential to quantify the uncertainty in bounding-box predictions for saf

safetyarxiv-cs-ai
4 Jun 2026
Safety

Multi-Agent Next-Best-View Optimization for Risk-Averse Planning

DGX agent

arXiv:2606.04158v1 Announce Type: new Abstract: Multi-agent Next-Best-View (NBV) selection for safe path planning in uncertain and unknown environments requires informative, safety-aware, and efficien

safetyarxiv-cs-ro
4 Jun 2026
Safety

Easy-to-Use Shielding for Reinforcement Learning

DGX agent

arXiv:2606.03804v1 Announce Type: new Abstract: Safe exploration is a key challenge in Reinforcement Learning (RL) that aims to prevent agents from making harmful decisions while exploring their envir

safetyarxiv-cs-lg
3 Jun 2026
Safety

Jailbreak Attack Initializations as Extractors of Compliance Directions

DGX agent

arXiv:2502.09755v4 Announce Type: replace-cross Abstract: Safety-aligned LLMs respond to prompts with either compliance or refusal, each corresponding to distinct directions in the model's activation

safetyarxiv-cs-lg
3 Jun 2026
Safety

Embedding Semantic Risk into Distance Fields and CBFs for Online Monocular Safe Control

DGX agent

arXiv:2606.01605v1 Announce Type: new Abstract: We propose an online monocular perception-to-control framework that embeds semantic risk into the distance field used by Control Barrier Function (CBF)-

safetyarxiv-cs-ro
2 Jun 2026
Safety

Multi-Objective Reinforcement Learning for Tactical Decision Making for Trucks in Highway Traffic

DGX agent

arXiv:2601.18783v2 Announce Type: replace-cross Abstract: Balancing safety, efficiency, and operational costs in highway driving poses a challenging decision-making problem for heavy-duty vehicles. A

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Predicted-Flow Control Barrier Functions for Real-Time Safe Optimal Control

DGX agent

arXiv:2606.00297v1 Announce Type: cross Abstract: Control barrier functions (CBFs) provide real-time safety guarantees through pointwise conditions on the state. However, synthesizing a valid CBF is d

model-releasesarxiv-cs-ro
2 Jun 2026
Safety

RedDebate: Safer Responses Through Multi-Agent Red Teaming Debates

DGX agent

arXiv:2506.11083v3 Announce Type: replace Abstract: We introduce RedDebate, a novel multi-agent debate framework that provides the foundation for Large Language Models (LLMs) to identify and mitigate

safetyarxiv-cs-cl
2 Jun 2026
Safety

Silent Failures in Physical AI: A Literature Review of Runtime Action Authorization for Autonomous Systems

DGX agent

arXiv:2606.00090v1 Announce Type: cross Abstract: Physical AI systems increasingly map multimodal observations, language instructions, and learned world representations into physically consequential a

safetyarxiv-cs-ai
2 Jun 2026
Safety

The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer

DGX agent

arXiv:2602.02557v2 Announce Type: replace-cross Abstract: Recent advances in end-to-end trained omni-models have substantially improved audio capabilities by strengthening text-audio modality alignmen

safetyarxiv-cs-ai
2 Jun 2026
Safety

From Out-of-Distribution Detection to Hallucination Detection: A Geometric View

DGX agent

arXiv:2602.07253v2 Announce Type: replace Abstract: Detecting hallucinations in large language models is a critical open problem with significant implications for safety and reliability. While existin

safetyarxiv-cs-ai
1 Jun 2026
Safety

A Review of Learning-Based Motion Planning: Toward a Data-Driven Optimal Control Approach

DGX agent

arXiv:2512.11944v2 Announce Type: replace-cross Abstract: Motion planning for autonomous driving (AD) faces a critical trade-off. While traditional rule-based pipelines offer verifiable safety and int

safetyarxiv-cs-ai
29 May 2026
Safety

Evolving and Detecting Multi-Turn Deception using Geometric Signatures

DGX agent

arXiv:2605.27671v1 Announce Type: cross Abstract: Safety defenses for large language models (LLMs) are typically trained and evaluated on single-turn prompts, yet real attacks often unfold as indirect

safetyarxiv-cs-lg
28 May 2026
Safety

High-Fidelity Industrial Crash Dynamics Prediction via Geometry-Aware Operator Learning with Memory-Efficient Low-Rank Attention

DGX agent

arXiv:2605.27758v1 Announce Type: cross Abstract: Automotive crashworthiness optimization remains a safety-critical challenge, requiring the management of large-scale nonlinear structural deformations

safetyarxiv-cs-ai
28 May 2026
Safety

Detecting Is Not Resolving: The Monitoring Control Gap in Retrieval Augmented LLMs

DGX agent

arXiv:2605.27157v1 Announce Type: new Abstract: Retrieval-augmented LLMs are deployed for tasks where evidence quality determines action safety, yet evaluation protocols assume that single-turn robust

safetyarxiv-cs-ai
27 May 2026
Safety

DBPnet: Damper Characteristics-Based Bayesian Physics-Informed Neural Network for Wheel Load Estimation

DGX agent

arXiv:2605.24860v1 Announce Type: cross Abstract: Advanced driver assistance systems (ADAS) play an important role in modern automotive intelligence, significantly enhancing vehicle safety and stabili

safetyarxiv-cs-ai
26 May 2026
Safety

From Knowledge to Inference: Formalizing Specialized Public Health Reasoning on GlobalHealthAtlas

DGX agent

arXiv:2602.00491v2 Announce Type: replace Abstract: Public health reasoning requires population level inference grounded in scientific evidence, expert consensus, and safety constraints. However, it r

safetyarxiv-cs-cl
26 May 2026
Safety

Empowering 9-1-1 Calltaking Training with Generative AI: Experiences and Lessons Learned

DGX agent

arXiv:2602.13241v2 Announce Type: replace-cross Abstract: Emergency call-takers form the first operational link in public safety response, handling over 240 million calls annually while facing a susta

safetyarxiv-cs-ai
25 May 2026
← Previous
1…1819202122…257
Next →