AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,237 results
11 May 2026

InvThink: Premortem Reasoning for Safer Language Models

SafetyDGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

7 May 2026

Brainrot: Deskilling and Addiction are Overlooked AI Risks

SafetyDGX agent

arXiv:2605.03512v1 Announce Type: cross Abstract: The scope of AI safety and alignment work in generative artificial intelligence (GenAI) has so far mostly been limited to harms related to: (a) discri

Practical validation of synthetic pre-crash scenarios

SafetyDGX agent

arXiv:2605.04564v1 Announce Type: new Abstract: The representativeness of synthetic pre-crash scenarios is crucial for assessing the safety impact of Driving Automation Systems through virtual simulat

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
5 May 2026

Lateral String Stability for Vehicle Platoons: Formulation, Definition, and Analysis

SafetyDGX agent

arXiv:2605.01731v1 Announce Type: new Abstract: Platooning of connected and automated vehicles provides significant benefits in terms of energy efficiency, traffic throughput, and, most critically, sa

30 Apr 2026

Risk Reporting for Developers' Internal AI Model Use

SafetyDGX agent

arXiv:2604.24966v1 Announce Type: cross Abstract: Frontier AI companies first deploy their most advanced models internally, for weeks or months of safety testing, evaluation, and iteration, before a p

28 Apr 2026

A Decoupled Human-in-the-Loop System for Controlled Autonomy in Agentic Workflows

SafetyDGX agent

arXiv:2604.23049v1 Announce Type: new Abstract: AI agents are increasingly deployed to execute tasks and make decisions within agentic workflows, introducing new requirements for safe and controlled a

CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning

SafetyDGX agent

arXiv:2604.23576v1 Announce Type: cross Abstract: Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning meth

Certified geometric robustness -- Super-DeepG

SafetyDGX agent

arXiv:2604.24379v1 Announce Type: new Abstract: Safety-critical applications are required to perform as expected in normal operations. Image processing functions are often required to be insensitive t

Early Warning of Intraoperative Adverse Events via Transformer-Driven Multi-Label Learning

SafetyDGX agent

arXiv:2603.05212v2 Announce Type: replace-cross Abstract: Early warning of intraoperative adverse events plays a vital role in reducing surgical risk and improving patient safety. While deep learning

Zoom In, Reason Out: Efficient Far-field Anomaly Detection in Expressway Surveillance Videos via Focused VLM Reasoning Guided by Bayesian Inference

SafetyDGX agent

arXiv:2604.23724v1 Announce Type: cross Abstract: Expressway video anomaly detection is essential for safety management. However, identifying anomalies across diverse scenes remains challenging, parti

27 Apr 2026

Improving Driver Drowsiness Detection via Personalized EAR/MAR Thresholds and CNN-Based Classification

SafetyDGX agent

arXiv:2604.22479v1 Announce Type: new Abstract: Driver drowsiness is a major cause of traffic accidents worldwide, posing a serious threat to public safety. Vision-based driver monitoring systems ofte

24 Apr 2026

How VLAs (Really) Work In Open-World Environments

SafetyDGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

The Economics of p(doom): Scenarios of Existential Risk and Economic Growth in the Age of Transformative AI

SafetyDGX agent

arXiv:2503.07341v2 Announce Type: replace-cross Abstract: Recent advances in artificial intelligence (AI) have led to a wide range of predictions about its long-term impact on humanity. A central focu

Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception

SafetyDGX agent

arXiv:2604.20895v1 Announce Type: cross Abstract: Safety and security are essential for the admission and acceptance of automated and autonomous vehicles. Deep neural networks (DNNs) are widely used f

23 Apr 2026

OnSiteVRU: A High-Resolution Trajectory Dataset for High-Density Vulnerable Road Users

SafetyDGX agent

arXiv:2503.23365v2 Announce Type: replace Abstract: With the acceleration of urbanization and the growth of transportation demands, the safety of vulnerable road users (VRUs, such as pedestrians and c

Physics-Enhanced Deep Learning for Proactive Thermal Runaway Forecasting in Li-Ion Batteries

SafetyDGX agent

arXiv:2604.20175v1 Announce Type: cross Abstract: Accurate prediction of thermal runaway in lithium-ion batteries is essential for ensuring the safety, efficiency, and reliability of modern energy sto

Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model

SafetyDGX agent

arXiv:2604.19838v1 Announce Type: new Abstract: Understanding how road users resolve space-sharing conflicts is important both for traffic safety and the safe deployment of autonomous vehicles. While

22 Apr 2026

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infra…

SafetyDGX agent

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infrastructure & AI software that underpins our Robotaxi & future

21 Apr 2026

Hybrid Spectro-Temporal Fusion Framework for Structural Health Monitoring

SafetyDGX agent

arXiv:2604.16589v1 Announce Type: new Abstract: Structural health monitoring plays a critical role in ensuring structural safety by analyzing vibration responses from engineering systems. This paper p

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

SafetyDGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

Safer Trajectory Planning with CBF-guided Diffusion Model for Unmanned Aerial Vehicles

SafetyDGX agent

arXiv:2604.17527v1 Announce Type: new Abstract: Safe and agile trajectory planning is essential for autonomous systems, especially during complex aerobatic maneuvers. Motivated by the recent success o

17 Apr 2026

Conformal Policy Control

SafetyDGX agent

arXiv:2603.02196v2 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm

16 Apr 2026

Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning

SafetyDGX agent

arXiv:2604.13891v1 Announce Type: new Abstract: Automated driving at unsignalized intersections is challenging due to complex multi-vehicle interactions and the need to balance safety and efficiency.

15 Apr 2026

4/5 We upgraded our original 3-line “be correct” prompt → a much more detailed prompt that enforced a hierarchy of constraints for correctne…

SafetyDGX agent

4/5 We upgraded our original 3-line “be correct” prompt → a much more detailed prompt that enforced a hierarchy of constraints for correctness, regression safety, and minimality. Basically, get the LL

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

Model ReleasesDGX agent

arXiv:2512.20798v4 Announce Type: replace Abstract: As autonomous AI agents are deployed in high-stakes environments, ensuring their safety has become a paramount concern. Existing safety benchmarks p

Cloud CISO Perspectives: How CISOs can pursue technical and cultural resilience (Q&A)

SafetyDGX agent

Welcome to the first Cloud CISO Perspectives for April 2026. Today, Thiébaut Meyer and Lia Wertheimer from Google Cloud’s Office of the CISO share Thiébaut’s conversation with Matt Rowe, chief securit

14 Apr 2026

Efficient Disruption of Criminal Networks through Multi-Objective Genetic Algorithms

SafetyDGX agent

arXiv:2604.09647v1 Announce Type: cross Abstract: Criminal networks, such as the Sicilian Mafia, pose substantial threats to public safety, national security, and economic stability. Outdated disrupti

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

SafetyDGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

10 Apr 2026

Designing Safe and Accountable GenAI as a Learning Companion with Women Banned from Formal Education

SafetyDGX agent

arXiv:2604.07253v1 Announce Type: cross Abstract: In gender-restrictive and surveilled contexts, where access to formal education may be restricted for women, pursuing education involves safety and pr

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

SafetyDGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

Leveraging Wireless Sensor Networks for Real-Time Monitoring and Control of Industrial Environments

SafetyDGX agent

arXiv:2510.13820v3 Announce Type: replace-cross Abstract: This research proposes an extensive technique for monitoring and controlling the industrial parameters using Internet of Things (IoT) technolo

11 Aug 2026

Autonomous Driving with Priority-Ordered STL Specifications Under Multimodal Uncertainty

SafetyDGX agent

arXiv:2606.20336v2 Announce Type: replace Abstract: Autonomous vehicles must plan trajectories that satisfy multiple requirements, such as safety, traffic-rule compliance, and passenger comfort. Howev

Pragmatic Attack Surface: Vulnerabilities of Implicit Context in Large Language Models

SafetyDGX agent

arXiv:2608.09551v1 Announce Type: new Abstract: In the era of large language models (LLMs), attackers often manipulate natural language to elicit unsafe or harmful outputs, creating a new natural lang

SAFE-CHEM: Uncertainty-Aware Policy Switching for Robust Robotic Chemistry

SafetyDGX agent

arXiv:2608.09303v1 Announce Type: cross Abstract: The deployment of autonomous robotic systems in chemistry laboratories is accelerating experimental workflows and providing the foundational data for

SafeSceneReason: A Multimodal Reasoning Benchmark Connecting Industrial Hazards with Accident Knowledge

Model ReleasesDGX agent

arXiv:2608.09230v1 Announce Type: new Abstract: Industrial-safety understanding requires more than detecting workers, equipment, and personal protective equipment. Models must also assess compliance,

SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding

SafetyDGX agent

arXiv:2512.09062v2 Announce Type: replace Abstract: Accurate 3D scene interpretation in active construction sites is essential for progress monitoring, safety assessment, and digital twin development.

Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM

SafetyDGX agent

arXiv:2511.18721v4 Announce Type: replace-cross Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict 'k-unstable' assumption that

10 Aug 2026

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “…

SafetyDGX agent

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “who distributes it to the most people”: Instead of fighting

7 Aug 2026

Safe Evolution with Circuit Anchors

SafetyDGX agent

arXiv:2608.05158v1 Announce Type: new Abstract: In biological evolution, unconstrained mutation can lead to catastrophic outcomes: organisms may evolve enhanced capabilities while losing essential fun

6 Aug 2026

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

SafetyDGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

SafeLand: Safe Autonomous Landing in Unknown Environments with Bayesian Semantic Mapping

SafetyDGX agent

arXiv:2603.17430v2 Announce Type: replace Abstract: Autonomous landing of uncrewed aerial vehicles (UAVs) in unknown, dynamic environments poses significant safety challenges, particularly near people

When Memory Lies: An Empirical Study of Spatial Memory Staleness in VLM Agents

SafetyDGX agent

arXiv:2608.04574v1 Announce Type: new Abstract: Memory-augmented VLM agents act on persistent spatial knowledge, yet that knowledge silently goes stale as the environment changes. We ask what happens

5 Aug 2026

Test-Time Scaling for Safe Text-Guided Image Generation via Intermediate Clean Estimates

SafetyDGX agent

arXiv:2608.03284v1 Announce Type: cross Abstract: Ensuring safety and policy compliance in text-to-image diffusion models remains a critical challenge, as benign or adversarial prompts can often elici

4 Aug 2026

Grounded Vision-Language Interpreter for Long-Horizon Bimanual Task and Motion Planning

SafetyDGX agent

arXiv:2506.03270v3 Announce Type: replace Abstract: While recent advances in vision-language models have accelerated language-guided robot planning, their black-box nature lacks the safety guarantees

The model takes moderation policy as a plain-language question and returns a calibrated score. Text and images — one interface. Read the ful…

SafetyDGX agent

Mistral AI has released Shieldstral, a 3‑billion‑parameter open‑weight model for content safety that can run locally on device. It interprets plain‑language moderation queries and returns calibrated s

3 Aug 2026

Beyond Component Testing: Validating Agentic AI Systems

SafetyDGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving

SafetyDGX agent

arXiv:2602.07339v2 Announce Type: replace Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denoising process introduces a latency bottleneck fo

TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text

SafetyDGX agent

arXiv:2607.28862v1 Announce Type: cross Abstract: The rapid development of Large Language Models (LLMs) has led to significant advances across a wide range of language tasks, while simultaneously rais

30 Jul 2026

Cohere has joined @NVIDIA alongside industry leaders in founding the Open Secure AI Alliance. Everybody should have the capability to keep t…

SafetyDGX agent

Cohere has joined @NVIDIA alongside industry leaders in founding the Open Secure AI Alliance. Everybody should have the capability to keep their infrastructure secure. Everybody deserves access to mod

29 Jul 2026

Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment

SafetyDGX agent

arXiv:2607.26034v1 Announce Type: new Abstract: Technological races create tension between speed and safety: actors may gain by moving faster than competitors, even when risky development is harmful.

Why Public Service AI Governance Frameworks Risk Failing in the Age of General-Purpose AI: Lessons from Policing

SafetyDGX agent

arXiv:2607.25648v1 Announce Type: cross Abstract: Public services face growing pressure to adopt artificial intelligence (AI) to close the gap between rising demand and falling resources. That pressur

28 Jul 2026

An Unofficial FastLAS Tutorial: A Programmer's Guide

SafetyDGX agent

arXiv:2607.23557v1 Announce Type: cross Abstract: FastLAS is a scalable system for Inductive Logic Programming (ILP): you give it some background knowledge, a language bias, and a set of examples, and

Synthetic Scenario Generation for Evaluation of Industry 4.0 Agents

SafetyDGX agent

arXiv:2607.22563v1 Announce Type: new Abstract: Industrial agent benchmarks require realistic evaluation scenarios that integrate telemetry, failure modes, maintenance records, and domain standards. H

Unifying Complementarity Constraints and Control Barrier Functions for Safe Whole-Body Robot Control

SafetyDGX agent

arXiv:2504.17647v2 Announce Type: replace Abstract: Safety-critical whole-body robot control demands reactive methods that ensure collision avoidance in real-time. Complementarity constraints and cont

24 Jul 2026

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models

SafetyDGX agent

arXiv:2607.20436v1 Announce Type: cross Abstract: Safety evaluations often assume that behavior observed during testing reflects behavior in ordinary use, but fine-tuning can break this assumption. A

16 Jul 2026

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

SafetyDGX agent

arXiv:2607.13172v1 Announce Type: new Abstract: We address the problem of safely training an agent policy and deploying a good and safe policy, in settings where the environment dynamics are unknown a

15 Jul 2026

APPLV: Adaptive Planner Parameter Learning from Vision-Language-Action Model

Model ReleasesDGX agent

arXiv:2603.08862v2 Announce Type: replace-cross Abstract: Autonomous navigation in highly constrained environments remains challenging for mobile robots. Classical navigation approaches offer safety a

Directional Constraints for Efficient Exploration in Safe Reinforcement Learning

SafetyDGX agent

arXiv:2607.12784v1 Announce Type: cross Abstract: Reinforcement Learning has revolutionized the landscape of robotic research, allowing robust learning of complex robotic skills in simulation. However

Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering

SafetyDGX agent

arXiv:2603.28583v2 Announce Type: replace-cross Abstract: Despite the success of Vision-Language Models (VLMs), misleading charts remain a significant challenge due to their deceptive visual structure

10 Jul 2026

INTENT: An LSTM Framework for Vehicle Intention Prediction in Intersection Scenarios with Comprehensive Ablation Analysis

SafetyDGX agent

arXiv:2607.08316v1 Announce Type: new Abstract: Vehicle intention prediction is a pivotal aspect in the agility and safety of autonomous vehicles in all driving scenarios; if genuine enhancement of au

← Previous
1…1516171819…238
Next →