AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Chai: Agentic Discovery of Cryptographic Misuse Vulnerabilities

DGX agent

arXiv:2606.26933v1 Announce Type: cross Abstract: AI-assisted vulnerability discovery has proven effective for bug classes like memory safety, where instrumentation confirms memory violations and effi

safetyarxiv-cs-ai
26 Jun 2026
Safety

HauntAttack: When Attack Follows Reasoning as a Shadow

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2506.07031v5 Announce Type: replace-cross Abstract: Emerging Large Reasoning Models (LRMs) consistently excel in mathematical and reasoning tasks, showcasing remarkable capabilities. However, th

safetyarxiv-cs-ai
26 Jun 2026
Safety

RecallRisk-BERT: A Multi-Task Framework for Post-Report Medical Device Recall Triage

DGX agent

arXiv:2606.27174v1 Announce Type: new Abstract: Medical device recalls are a critical regulatory mechanism for protecting patient safety. The growing volume of FDA recall records presents challenges i

safetyarxiv-cs-lg
26 Jun 2026
Model Releases

Falcon: Functional Assembly and Language for Compositional Reasoning in X-ray

DGX agent

arXiv:2606.25701v1 Announce Type: new Abstract: Conventional vision-language models are largely object-centric, focusing on detecting and describing individual entities. In safety-critical X-ray bagga

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

The Tatoxa System for Text Detoxification in Low-Resource Languages: The Case of Tatar

DGX agent

arXiv:2606.26015v1 Announce Type: new Abstract: Text detoxification, the automated detection and mitigation of abusive and harmful content, is essential for ensuring the safety of online communities a

safetyarxiv-cs-cl
25 Jun 2026
Safety

Towards Understanding The Calibration Benefits of Sharpness-Aware Minimization

DGX agent

arXiv:2505.23866v2 Announce Type: replace Abstract: Deep neural networks have been increasingly used in safety-critical applications such as medical diagnosis and autonomous driving. However, many stu

safetyarxiv-cs-lg
25 Jun 2026
Safety

DDStereo: Efficient Dual Decoder Transformers for Stereo 3D Road Anomaly Detection

DGX agent

arXiv:2606.24805v1 Announce Type: new Abstract: Stereo-based 3D object detection still faces two critical safety challenges: real-time performance and open-set generalization. Existing stereo 3D metho

safetyarxiv-cs-cv
24 Jun 2026
Safety

One Year Later...The Harms Persist, But So Do We!

DGX agent

arXiv:2606.23884v1 Announce Type: cross Abstract: General-purpose large language models (LLMs) are increasingly used for mental health-related conversations, yet safety safeguards remain inadequate an

safetyarxiv-cs-ai
24 Jun 2026
Safety

An LLM-Explainable DRL Framework for Passenger-Directed Autonomous Driving

DGX agent

arXiv:2606.20640v1 Announce Type: cross Abstract: Autonomous vehicles offer the potential for safer and more efficient mobility, yet public trust remains limited due to the lack of transparency in the

safetyarxiv-cs-lg
23 Jun 2026
Safety

BayesFP: Posterior Estimation for Flow-Based Policies via Feynman-Kac Sampling

DGX agent

arXiv:2606.21014v1 Announce Type: new Abstract: Robots must generate trajectories that remain faithful to learned expert behavior while satisfying safety constraints and task-specific objectives speci

safetyarxiv-cs-ro
23 Jun 2026
Safety

Conflict-Aware Switching for CBF-CLF-Based Multi-Goal Navigation

DGX agent

arXiv:2606.21577v1 Announce Type: new Abstract: Quadratic programs (QPs) using Control Barrier Functions (CBFs) and Control Lyapunov Functions (CLFs) are widely used for safe control in reach-and-avoi

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

CRAX: Fast Safe Reinforcement Learning Benchmarking

DGX agent

arXiv:2606.20376v2 Announce Type: replace Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While bench

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

DBT-Bleed: Dual-Branch Temporal Modeling with Key-Frame Selection for Surgical Bleeding Detection

DGX agent

arXiv:2606.22829v1 Announce Type: new Abstract: Intraoperative Adverse Events (IAEs) detection is critical for improving surgical safety, with bleeding being among the most frequent events across many

safetyarxiv-cs-cv
23 Jun 2026
Safety

From Driving Videos to Simulatable Scenarios

DGX agent

arXiv:2606.21993v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) face driving scenarios ranging from routine traffic to rare events. To assess safety it is crucial to reproduce these scenar

safetyarxiv-cs-cv
23 Jun 2026
Safety

NegAS: Negative Label Guided Attention and Scoring for Out-of-Distribution Object Detection with Vision-Language Models

DGX agent

arXiv:2606.22537v1 Announce Type: new Abstract: Out-of-Distribution (OOD) detection is essential for ensuring the robustness and reliability of object detection systems deployed in safety-critical app

safetyarxiv-cs-cv
23 Jun 2026
Safety

Neural Architecture Distributions: A New Paradigm for Stochastic Segmentation

DGX agent

arXiv:2606.21061v1 Announce Type: new Abstract: Stochastic segmentation seeks to represent multiple plausible masks for a single image, which is essential in safety- and quality-critical applications

safetyarxiv-cs-cv
23 Jun 2026
Safety

Predicting cognitive load in immersive driving scenarios with a hybrid CNN-RNN model

DGX agent

arXiv:2408.06350v2 Announce Type: replace-cross Abstract: One debatable issue in traffic safety research is that cognitive load from sec-ondary tasks reduces primary task performance, such as driving.

safetyarxiv-cs-lg
23 Jun 2026
Safety

Safe Few-Step Generation via Velocity Editing

DGX agent

arXiv:2606.23267v1 Announce Type: new Abstract: Flow matching has recently emerged as a strong paradigm for state-of-the-art text-to-image (T2I) generation, enabling high-quality generation with a sma

safetyarxiv-cs-cv
23 Jun 2026
Safety

SPiralRoll: A Novel Adjustable-Stiffness Underactuated 3-DoF Joint with Torsion Springs for Rolling Robots

DGX agent

arXiv:2606.22443v1 Announce Type: new Abstract: Compliant mechanisms are important in robotics because they can improve adaptability, safety, and energy efficiency while reducing hardware complexity.

safetyarxiv-cs-ro
23 Jun 2026
Safety

AutoMine Solution for AV2 2026 Scenario Mining Challenge

DGX agent

arXiv:2606.11874v1 Announce Type: new Abstract: With the development of autonomous driving systems, mining high-value, safety-critical, and planning-relevant scenarios from large-scale driving logs ha

safetyarxiv-cs-ai
11 Jun 2026
Safety

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

DGX agent

arXiv:2606.11648v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean

safetyarxiv-cs-cl
11 Jun 2026
Safety

Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers

DGX agent

arXiv:2606.12025v1 Announce Type: new Abstract: Finite element (FE) modeling of safety-critical infrastructure such as bridge barriers requires high-fidelity nonlinear dynamic analysis, yet the curren

safetyarxiv-cs-ai
11 Jun 2026
Safety

JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization

DGX agent

arXiv:2606.11425v1 Announce Type: cross Abstract: Jailbreak attacks expose persistent safety weaknesses in large language models (LLMs), but existing stateless single-turn methods face a trade-off: ha

safetyarxiv-cs-ai
11 Jun 2026
Safety

One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection

DGX agent

arXiv:2606.11202v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in domina

safetyarxiv-cs-cl
11 Jun 2026
Safety

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models

DGX agent

arXiv:2606.11266v1 Announce Type: new Abstract: The cost signal that constrained-RL algorithms optimize against is almost always reactive: the simulator emits a non-zero cost only after a collision ha

safetyarxiv-cs-lg
11 Jun 2026
Safety

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

DGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

safetyarxiv-cs-ai
10 Jun 2026
Safety

Automatic Labelling for Low-Light Pedestrian Detection

DGX agent

arXiv:2507.02513v4 Announce Type: replace Abstract: Pedestrian detection in RGB images is a key task in pedestrian safety, as the most common sensor in autonomous vehicles and advanced driver assistan

safetyarxiv-cs-cv
10 Jun 2026
Safety

On the Controllability-Fidelity Frontier in Diffusion Editing

DGX agent

arXiv:2606.09901v1 Announce Type: cross Abstract: Diffusion-based generative models enable powerful image editing capabilities, but achieving precise control while maintaining fidelity and safety rema

safetyarxiv-cs-cv
10 Jun 2026
Safety

A VideoMAE-v2 Approach to Zero-Shot Traffic Accident Anticipation

DGX agent

arXiv:2606.09542v1 Announce Type: new Abstract: Traffic accident anticipation -- predicting the likelihood of an imminent collision at every frame of a dashcam video -- is safety-critical yet difficul

safetyarxiv-cs-cv
9 Jun 2026
Safety

Automated Framework to Evaluate and Harden LLM System Instructions against Encoding Attacks

DGX agent

arXiv:2604.01039v2 Announce Type: replace-cross Abstract: System Instructions in Large Language Models (LLMs) are commonly used to enforce safety policies, define agent behavior, and protect sensitive

safetyarxiv-cs-ai
9 Jun 2026
Safety

Code Is More Than Text: Uncertainty Estimation for Code Generation

DGX agent

arXiv:2606.09577v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as code generators, where silently wrong programs pose real safety and reliability risks. Relia

safetyarxiv-cs-lg
9 Jun 2026
Safety

Distant Object Localisation from Noisy Image Segmentation Sequences

DGX agent

arXiv:2509.20906v3 Announce Type: replace Abstract: 3D object localisation based on a sequence of camera measurements is essential for safety-critical surveillance tasks, such as drone-based wildfire

safetyarxiv-cs-cv
9 Jun 2026
Safety

Human-Centered Benchmarking of Driver Monitoring Models

DGX agent

arXiv:2606.08123v1 Announce Type: cross Abstract: Vision-based driver monitoring systems are increasingly deployed in safety-critical intelligent transportation settings, yet they are almost always co

safetyarxiv-cs-ai
9 Jun 2026
Safety

Hyperspectral Smoke Segmentation via Mixture of Prototypes

DGX agent

arXiv:2602.10858v2 Announce Type: replace Abstract: Smoke segmentation is critical for wildfire management and industrial safety applications. Traditional visible-light-based methods face limitations

safetyarxiv-cs-cv
9 Jun 2026
Safety

Large Language Models Should Learn Personalized Rather Than Aggregated Human Preferences

DGX agent

arXiv:2606.07629v1 Announce Type: cross Abstract: Current approaches to aligning large language models (LLMs) aggregate diverse human preferences into a single reward signal, effectively optimizing fo

safetyarxiv-cs-ai
9 Jun 2026
Safety

LUNA-AD: Lightweight Uncertainty-Aware Language Model with Lifelong Learning for Autonomous Driving

DGX agent

arXiv:2606.08470v1 Announce Type: new Abstract: While large language models (LLMs) offer promising reasoning capabilities, their integration into safety-critical driving systems is hindered by limited

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

DGX agent

arXiv:2606.07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge. Whi

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

On-the-fly hand-eye calibration for the da Vinci surgical robot

DGX agent

arXiv:2601.14871v2 Announce Type: replace Abstract: In Robot-Assisted Minimally Invasive Surgery (RMIS), accurate tool localization is crucial to ensure patient safety and successful task execution. H

safetyarxiv-cs-ro
9 Jun 2026
Safety

QDS-SNN: Energy-efficient Quantum Deeply-Supervised Spiking Neural Network Algorithm for Traffic Sign Recognition

DGX agent

arXiv:2606.07657v1 Announce Type: cross Abstract: Traffic sign recognition is crucial for intelligent transportation and autonomous driving, as it can improve driving efficiency and ensure road safety

safetyarxiv-cs-lg
9 Jun 2026
Safety

Scaling Neural Network Verification with Tensor Parallelism and Fully Sharded Data Parallelism

DGX agent

arXiv:2606.09377v1 Announce Type: cross Abstract: Formal neural network verification -- proving that a network satisfies safety properties for all inputs in a specified domain -- is bounded in practic

safetyarxiv-cs-ai
9 Jun 2026
Safety

State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space

DGX agent

arXiv:2601.04266v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are widely deployed in safety-critical embodied AI applications such as robotics. However, their complex m

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Confidence Trap: Calibration Attacks for Graph Neural Networks

DGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

safetyarxiv-cs-ai
9 Jun 2026
Safety

Vessel Traffic Flow Prediction on Sparse Data via Spatio-Temporal Graph Neural Networks with a Learnable Tweedie Head

DGX agent

arXiv:2606.07694v1 Announce Type: new Abstract: Accurate vessel traffic flow prediction is crucial for smart port operations and navigational safety. However, maritime traffic flow data are often high

safetyarxiv-cs-lg
9 Jun 2026
Safety

Latent-space Attacks for Refusal Evasion in Language Models

DGX agent

arXiv:2605.21706v2 Announce Type: replace Abstract: Safety-aligned language models are trained to refuse harmful requests, yet refusal behavior can be suppressed by steering their internal representat

safetyarxiv-cs-ai
8 Jun 2026
Safety

Multi-Scale Feature Attention Network for Polymer Classification using THz Dual-Comb Spectroscopy

DGX agent

arXiv:2606.06554v1 Announce Type: cross Abstract: Reliable polymer identification is essential for ensuring the quality and safety of recycled plastics, yet conventional sorting and spectroscopic tech

safetyarxiv-cs-ai
8 Jun 2026
Safety

Residual-Controlled Multiplier Learning for Stochastic Constrained Decision-Making

DGX agent

arXiv:2606.07088v1 Announce Type: new Abstract: Stochastic constrained decision-making requires optimizing performance objectives while enforcing statistical requirements such as safety or fairness. H

safetyarxiv-cs-lg
8 Jun 2026
Safety

Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach

DGX agent

arXiv:2510.09041v3 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has demonstrated remarkable success in developing autonomous driving policies. However, its vulnerability to

safetyarxiv-cs-ai
8 Jun 2026
Safety

From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

DGX agent

arXiv:2606.06223v1 Announce Type: new Abstract: Language-model agents act through repeated cycles of observation, reasoning, and action selection, making safety monitoring depend on both internal mode

safetyarxiv-cs-ai
6 Jun 2026
← Previous
1…2324252627…257
Next →