AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,349 results
Safety

BayesFP: Posterior Estimation for Flow-Based Policies via Feynman-Kac Sampling

DGX agent

arXiv:2606.21014v1 Announce Type: new Abstract: Robots must generate trajectories that remain faithful to learned expert behavior while satisfying safety constraints and task-specific objectives speci

safetyarxiv-cs-ro
23 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Conflict-Aware Switching for CBF-CLF-Based Multi-Goal Navigation

DGX agent

arXiv:2606.21577v1 Announce Type: new Abstract: Quadratic programs (QPs) using Control Barrier Functions (CBFs) and Control Lyapunov Functions (CLFs) are widely used for safe control in reach-and-avoi

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

CRAX: Fast Safe Reinforcement Learning Benchmarking

DGX agent

arXiv:2606.20376v2 Announce Type: replace Abstract: Safety is a core concern for deploying reinforcement learning (RL) agents in real-world domains such as robotics and autonomous driving. While bench

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

DBT-Bleed: Dual-Branch Temporal Modeling with Key-Frame Selection for Surgical Bleeding Detection

DGX agent

arXiv:2606.22829v1 Announce Type: new Abstract: Intraoperative Adverse Events (IAEs) detection is critical for improving surgical safety, with bleeding being among the most frequent events across many

safetyarxiv-cs-cv
23 Jun 2026
Safety

From Driving Videos to Simulatable Scenarios

DGX agent

arXiv:2606.21993v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) face driving scenarios ranging from routine traffic to rare events. To assess safety it is crucial to reproduce these scenar

safetyarxiv-cs-cv
23 Jun 2026
Safety

My biggest NY-12 competitors aren't on the ballot today: Trump's AI oligarchs. If they can defeat me—the author of the nation's strongest AI…

DGX agent

My biggest NY-12 competitors aren't on the ballot today: Trump's AI oligarchs. If they can defeat me—the author of the nation's strongest AI safety bill in one of the nation's bluest districts—they ca

safetygary-marcus--x
23 Jun 2026
Safety

NegAS: Negative Label Guided Attention and Scoring for Out-of-Distribution Object Detection with Vision-Language Models

DGX agent

arXiv:2606.22537v1 Announce Type: new Abstract: Out-of-Distribution (OOD) detection is essential for ensuring the robustness and reliability of object detection systems deployed in safety-critical app

safetyarxiv-cs-cv
23 Jun 2026
Safety

Neural Architecture Distributions: A New Paradigm for Stochastic Segmentation

DGX agent

arXiv:2606.21061v1 Announce Type: new Abstract: Stochastic segmentation seeks to represent multiple plausible masks for a single image, which is essential in safety- and quality-critical applications

safetyarxiv-cs-cv
23 Jun 2026
Safety

Predicting cognitive load in immersive driving scenarios with a hybrid CNN-RNN model

DGX agent

arXiv:2408.06350v2 Announce Type: replace-cross Abstract: One debatable issue in traffic safety research is that cognitive load from sec-ondary tasks reduces primary task performance, such as driving.

safetyarxiv-cs-lg
23 Jun 2026
Safety

Safe Few-Step Generation via Velocity Editing

DGX agent

arXiv:2606.23267v1 Announce Type: new Abstract: Flow matching has recently emerged as a strong paradigm for state-of-the-art text-to-image (T2I) generation, enabling high-quality generation with a sma

safetyarxiv-cs-cv
23 Jun 2026
Safety

SPiralRoll: A Novel Adjustable-Stiffness Underactuated 3-DoF Joint with Torsion Springs for Rolling Robots

DGX agent

arXiv:2606.22443v1 Announce Type: new Abstract: Compliant mechanisms are important in robotics because they can improve adaptability, safety, and energy efficiency while reducing hardware complexity.

safetyarxiv-cs-ro
23 Jun 2026
Safety

AutoMine Solution for AV2 2026 Scenario Mining Challenge

DGX agent

arXiv:2606.11874v1 Announce Type: new Abstract: With the development of autonomous driving systems, mining high-value, safety-critical, and planning-relevant scenarios from large-scale driving logs ha

safetyarxiv-cs-ai
11 Jun 2026
Safety

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

DGX agent

arXiv:2606.11648v1 Announce Type: cross Abstract: Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean

safetyarxiv-cs-cl
11 Jun 2026
Safety

Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers

DGX agent

arXiv:2606.12025v1 Announce Type: new Abstract: Finite element (FE) modeling of safety-critical infrastructure such as bridge barriers requires high-fidelity nonlinear dynamic analysis, yet the curren

safetyarxiv-cs-ai
11 Jun 2026
Safety

JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization

DGX agent

arXiv:2606.11425v1 Announce Type: cross Abstract: Jailbreak attacks expose persistent safety weaknesses in large language models (LLMs), but existing stateless single-turn methods face a trade-off: ha

safetyarxiv-cs-ai
11 Jun 2026
Safety

One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection

DGX agent

arXiv:2606.11202v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in domina

safetyarxiv-cs-cl
11 Jun 2026
Safety

Seeing Before Colliding: Anticipatory Safe RL with Frozen Vision-Language Models

DGX agent

arXiv:2606.11266v1 Announce Type: new Abstract: The cost signal that constrained-RL algorithms optimize against is almost always reactive: the simulator emits a non-zero cost only after a collision ha

safetyarxiv-cs-lg
11 Jun 2026
Safety

A Reliable Fault Diagnosis Method Based on Belief Rule Base Consider Robustness Analysis

DGX agent

arXiv:2606.10500v1 Announce Type: new Abstract: In equipment operation, the implementation of fault diagnosis is essential to ensure the continuity and safety of production equipment, improve operatio

safetyarxiv-cs-ai
10 Jun 2026
Safety

Anthropic and OpenAI did not call for a pause. Read the wording: 'good for the world to have the option', 'possible' to slow down 'when need…

DGX agent

Anthropic and OpenAI did not call for a pause. Read the wording: 'good for the world to have the option', 'possible' to slow down 'when needed'. This is how they signal safety to one audience, acceler

safetyconnor-leahy--x
10 Jun 2026
Safety

Anthropic, if they really believe what they say, should show some leadership:

DGX agent

Anthropic, if they really believe what they say, should show some leadership: 🔔 IF Anthropic is for real about safety, the should pause, for one month, and show leadership. If they won’t pause, even b

safetygary-marcus--x
10 Jun 2026
Safety

Automatic Labelling for Low-Light Pedestrian Detection

DGX agent

arXiv:2507.02513v4 Announce Type: replace Abstract: Pedestrian detection in RGB images is a key task in pedestrian safety, as the most common sensor in autonomous vehicles and advanced driver assistan

safetyarxiv-cs-cv
10 Jun 2026
Safety

On the Controllability-Fidelity Frontier in Diffusion Editing

DGX agent

arXiv:2606.09901v1 Announce Type: cross Abstract: Diffusion-based generative models enable powerful image editing capabilities, but achieving precise control while maintaining fidelity and safety rema

safetyarxiv-cs-cv
10 Jun 2026
Safety

A VideoMAE-v2 Approach to Zero-Shot Traffic Accident Anticipation

DGX agent

arXiv:2606.09542v1 Announce Type: new Abstract: Traffic accident anticipation -- predicting the likelihood of an imminent collision at every frame of a dashcam video -- is safety-critical yet difficul

safetyarxiv-cs-cv
9 Jun 2026
Safety

Automated Framework to Evaluate and Harden LLM System Instructions against Encoding Attacks

DGX agent

arXiv:2604.01039v2 Announce Type: replace-cross Abstract: System Instructions in Large Language Models (LLMs) are commonly used to enforce safety policies, define agent behavior, and protect sensitive

safetyarxiv-cs-ai
9 Jun 2026
Safety

Code Is More Than Text: Uncertainty Estimation for Code Generation

DGX agent

arXiv:2606.09577v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as code generators, where silently wrong programs pose real safety and reliability risks. Relia

safetyarxiv-cs-lg
9 Jun 2026
Safety

Distant Object Localisation from Noisy Image Segmentation Sequences

DGX agent

arXiv:2509.20906v3 Announce Type: replace Abstract: 3D object localisation based on a sequence of camera measurements is essential for safety-critical surveillance tasks, such as drone-based wildfire

safetyarxiv-cs-cv
9 Jun 2026
Safety

Human-Centered Benchmarking of Driver Monitoring Models

DGX agent

arXiv:2606.08123v1 Announce Type: cross Abstract: Vision-based driver monitoring systems are increasingly deployed in safety-critical intelligent transportation settings, yet they are almost always co

safetyarxiv-cs-ai
9 Jun 2026
Safety

Hyperspectral Smoke Segmentation via Mixture of Prototypes

DGX agent

arXiv:2602.10858v2 Announce Type: replace Abstract: Smoke segmentation is critical for wildfire management and industrial safety applications. Traditional visible-light-based methods face limitations

safetyarxiv-cs-cv
9 Jun 2026
Safety

Large Language Models Should Learn Personalized Rather Than Aggregated Human Preferences

DGX agent

arXiv:2606.07629v1 Announce Type: cross Abstract: Current approaches to aligning large language models (LLMs) aggregate diverse human preferences into a single reward signal, effectively optimizing fo

safetyarxiv-cs-ai
9 Jun 2026
Safety

LUNA-AD: Lightweight Uncertainty-Aware Language Model with Lifelong Learning for Autonomous Driving

DGX agent

arXiv:2606.08470v1 Announce Type: new Abstract: While large language models (LLMs) offer promising reasoning capabilities, their integration into safety-critical driving systems is hindered by limited

safetyarxiv-cs-ro
9 Jun 2026
Model Releases

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

DGX agent

arXiv:2606.07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge. Whi

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

On-the-fly hand-eye calibration for the da Vinci surgical robot

DGX agent

arXiv:2601.14871v2 Announce Type: replace Abstract: In Robot-Assisted Minimally Invasive Surgery (RMIS), accurate tool localization is crucial to ensure patient safety and successful task execution. H

safetyarxiv-cs-ro
9 Jun 2026
Safety

QDS-SNN: Energy-efficient Quantum Deeply-Supervised Spiking Neural Network Algorithm for Traffic Sign Recognition

DGX agent

arXiv:2606.07657v1 Announce Type: cross Abstract: Traffic sign recognition is crucial for intelligent transportation and autonomous driving, as it can improve driving efficiency and ensure road safety

safetyarxiv-cs-lg
9 Jun 2026
Safety

Scaling Neural Network Verification with Tensor Parallelism and Fully Sharded Data Parallelism

DGX agent

arXiv:2606.09377v1 Announce Type: cross Abstract: Formal neural network verification -- proving that a network satisfies safety properties for all inputs in a specified domain -- is bounded in practic

safetyarxiv-cs-ai
9 Jun 2026
Safety

State Backdoor: Towards Stealthy Real-world Poisoning Attack on Vision-Language-Action Model in State Space

DGX agent

arXiv:2601.04266v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are widely deployed in safety-critical embodied AI applications such as robotics. However, their complex m

safetyarxiv-cs-lg
9 Jun 2026
Safety

The Confidence Trap: Calibration Attacks for Graph Neural Networks

DGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

safetyarxiv-cs-ai
9 Jun 2026
Safety

Vessel Traffic Flow Prediction on Sparse Data via Spatio-Temporal Graph Neural Networks with a Learnable Tweedie Head

DGX agent

arXiv:2606.07694v1 Announce Type: new Abstract: Accurate vessel traffic flow prediction is crucial for smart port operations and navigational safety. However, maritime traffic flow data are often high

safetyarxiv-cs-lg
9 Jun 2026
Safety

Built to benefit everyone: our plan

DGX agent

OpenAI outlines its mission and strategic approach to developing artificial intelligence technologies designed to benefit humanity broadly. The plan likely covers OpenAI's commitments to safety resear

safetyopenai
8 Jun 2026
Safety

Latent-space Attacks for Refusal Evasion in Language Models

DGX agent

arXiv:2605.21706v2 Announce Type: replace Abstract: Safety-aligned language models are trained to refuse harmful requests, yet refusal behavior can be suppressed by steering their internal representat

safetyarxiv-cs-ai
8 Jun 2026
Safety

Multi-Scale Feature Attention Network for Polymer Classification using THz Dual-Comb Spectroscopy

DGX agent

arXiv:2606.06554v1 Announce Type: cross Abstract: Reliable polymer identification is essential for ensuring the quality and safety of recycled plastics, yet conventional sorting and spectroscopic tech

safetyarxiv-cs-ai
8 Jun 2026
Safety

Residual-Controlled Multiplier Learning for Stochastic Constrained Decision-Making

DGX agent

arXiv:2606.07088v1 Announce Type: new Abstract: Stochastic constrained decision-making requires optimizing performance objectives while enforcing statistical requirements such as safety or fairness. H

safetyarxiv-cs-lg
8 Jun 2026
Safety

Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach

DGX agent

arXiv:2510.09041v3 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has demonstrated remarkable success in developing autonomous driving policies. However, its vulnerability to

safetyarxiv-cs-ai
8 Jun 2026
Safety

From Reward-Hack Activations to Agentic Risk States: Context-Calibrated Mechanistic Monitoring in LLM Agents

DGX agent

arXiv:2606.06223v1 Announce Type: new Abstract: Language-model agents act through repeated cycles of observation, reasoning, and action selection, making safety monitoring depend on both internal mode

safetyarxiv-cs-ai
6 Jun 2026
Safety

Towards Healthy Evolution: Exploring the Role and Mechanisms of Human-Agent Interaction in Self-Evolving Systems

DGX agent

arXiv:2606.06114v1 Announce Type: new Abstract: Self-evolving agents improve through continual self-play and self-generated learning signals, but autonomous evolution can also cause capability degrada

safetyarxiv-cs-ai
6 Jun 2026
Safety

Real-Time Threat Detection from Surveillance Cameras using Machine Learning

DGX agent

arXiv:2606.05708v1 Announce Type: new Abstract: Ensuring public safety in densely populated urban environments remains a critical challenge, necessitating the deployment of intelligent and automated v

safetyarxiv-cs-cv
5 Jun 2026
Safety

Expert-Aware Refusal Steering

DGX agent

arXiv:2606.04160v1 Announce Type: new Abstract: Safety alignment in instruction-tuned large language models (LLMs) depends on a model's ability to reliably refuse to respond to harmful or disallowed r

safetyarxiv-cs-cl
4 Jun 2026
Safety

Measuring Model Robustness via Fisher Information: Spectral Bounds, Theoretical Guarantees, and Practical Algorithms

DGX agent

arXiv:2606.04767v1 Announce Type: cross Abstract: The robustness of deep neural networks is crucial for safety-critical deployments, yet existing evaluation methods are often attack-dependent and lack

safetyarxiv-cs-cv
4 Jun 2026
Safety

Testing Neural Networks via Bayesian-Guided Exploration of Decision Landscapes

DGX agent

arXiv:2606.04314v1 Announce Type: new Abstract: As neural networks are increasingly deployed in safety-critical domains, testing is essential to evaluate and improve their reliability. Existing testin

safetyarxiv-cs-lg
4 Jun 2026
← Previous
1…2627282930…299
Next →