AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlog
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,237 results
Safety

InvThink: Premortem Reasoning for Safer Language Models

DGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

safetyarxiv-cs-ai
11 May 2026
Safety

Brainrot: Deskilling and Addiction are Overlooked AI Risks

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.03512v1 Announce Type: cross Abstract: The scope of AI safety and alignment work in generative artificial intelligence (GenAI) has so far mostly been limited to harms related to: (a) discri

safetyarxiv-cs-ai
7 May 2026
Safety

Practical validation of synthetic pre-crash scenarios

DGX agent

arXiv:2605.04564v1 Announce Type: new Abstract: The representativeness of synthetic pre-crash scenarios is crucial for assessing the safety impact of Driving Automation Systems through virtual simulat

safetyarxiv-cs-ro
7 May 2026
Safety

Lateral String Stability for Vehicle Platoons: Formulation, Definition, and Analysis

DGX agent

arXiv:2605.01731v1 Announce Type: new Abstract: Platooning of connected and automated vehicles provides significant benefits in terms of energy efficiency, traffic throughput, and, most critically, sa

safetyarxiv-cs-ro
5 May 2026
Safety

Risk Reporting for Developers' Internal AI Model Use

DGX agent

arXiv:2604.24966v1 Announce Type: cross Abstract: Frontier AI companies first deploy their most advanced models internally, for weeks or months of safety testing, evaluation, and iteration, before a p

safetyarxiv-cs-ai
30 Apr 2026
Safety

A Decoupled Human-in-the-Loop System for Controlled Autonomy in Agentic Workflows

DGX agent

arXiv:2604.23049v1 Announce Type: new Abstract: AI agents are increasingly deployed to execute tasks and make decisions within agentic workflows, introducing new requirements for safe and controlled a

safetyarxiv-cs-ai
28 Apr 2026
Safety

CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning

DGX agent

arXiv:2604.23576v1 Announce Type: cross Abstract: Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning meth

safetyarxiv-cs-ai
28 Apr 2026
Safety

Certified geometric robustness -- Super-DeepG

DGX agent

arXiv:2604.24379v1 Announce Type: new Abstract: Safety-critical applications are required to perform as expected in normal operations. Image processing functions are often required to be insensitive t

safetyarxiv-cs-ai
28 Apr 2026
Safety

Early Warning of Intraoperative Adverse Events via Transformer-Driven Multi-Label Learning

DGX agent

arXiv:2603.05212v2 Announce Type: replace-cross Abstract: Early warning of intraoperative adverse events plays a vital role in reducing surgical risk and improving patient safety. While deep learning

safetyarxiv-cs-ai
28 Apr 2026
Safety

Zoom In, Reason Out: Efficient Far-field Anomaly Detection in Expressway Surveillance Videos via Focused VLM Reasoning Guided by Bayesian Inference

DGX agent

arXiv:2604.23724v1 Announce Type: cross Abstract: Expressway video anomaly detection is essential for safety management. However, identifying anomalies across diverse scenes remains challenging, parti

safetyarxiv-cs-ai
28 Apr 2026
Safety

Improving Driver Drowsiness Detection via Personalized EAR/MAR Thresholds and CNN-Based Classification

DGX agent

arXiv:2604.22479v1 Announce Type: new Abstract: Driver drowsiness is a major cause of traffic accidents worldwide, posing a serious threat to public safety. Vision-based driver monitoring systems ofte

safetyarxiv-cs-cv
27 Apr 2026
Safety

How VLAs (Really) Work In Open-World Environments

DGX agent

arXiv:2604.21192v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have been extensively used in robotics applications, achieving great success in various manipulation problems. Mo

safetyarxiv-cs-ai
24 Apr 2026
Safety

The Economics of p(doom): Scenarios of Existential Risk and Economic Growth in the Age of Transformative AI

DGX agent

arXiv:2503.07341v2 Announce Type: replace-cross Abstract: Recent advances in artificial intelligence (AI) have led to a wide range of predictions about its long-term impact on humanity. A central focu

safetyarxiv-cs-ai
24 Apr 2026
Safety

Towards a Systematic Risk Assessment of Deep Neural Network Limitations in Autonomous Driving Perception

DGX agent

arXiv:2604.20895v1 Announce Type: cross Abstract: Safety and security are essential for the admission and acceptance of automated and autonomous vehicles. Deep neural networks (DNNs) are widely used f

safetyarxiv-cs-lg
24 Apr 2026
Safety

OnSiteVRU: A High-Resolution Trajectory Dataset for High-Density Vulnerable Road Users

DGX agent

arXiv:2503.23365v2 Announce Type: replace Abstract: With the acceleration of urbanization and the growth of transportation demands, the safety of vulnerable road users (VRUs, such as pedestrians and c

safetyarxiv-cs-cv
23 Apr 2026
Safety

Physics-Enhanced Deep Learning for Proactive Thermal Runaway Forecasting in Li-Ion Batteries

DGX agent

arXiv:2604.20175v1 Announce Type: cross Abstract: Accurate prediction of thermal runaway in lithium-ion batteries is essential for ensuring the safety, efficiency, and reliability of modern energy sto

safetyarxiv-cs-ai
23 Apr 2026
Safety

Resolving space-sharing conflicts in road user interactions through uncertainty reduction: An active inference-based computational model

DGX agent

arXiv:2604.19838v1 Announce Type: new Abstract: Understanding how road users resolve space-sharing conflicts is important both for traffic safety and the safe deployment of autonomous vehicles. While

safetyarxiv-cs-ai
23 Apr 2026
Safety

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infra…

DGX agent

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infrastructure & AI software that underpins our Robotaxi & future

safetyelon-musk--x
22 Apr 2026
Safety

Hybrid Spectro-Temporal Fusion Framework for Structural Health Monitoring

DGX agent

arXiv:2604.16589v1 Announce Type: new Abstract: Structural health monitoring plays a critical role in ensuring structural safety by analyzing vibration responses from engineering systems. This paper p

safetyarxiv-cs-lg
21 Apr 2026
Safety

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

DGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

safetyarxiv-cs-cv
21 Apr 2026
Safety

Safer Trajectory Planning with CBF-guided Diffusion Model for Unmanned Aerial Vehicles

DGX agent

arXiv:2604.17527v1 Announce Type: new Abstract: Safe and agile trajectory planning is essential for autonomous systems, especially during complex aerobatic maneuvers. Motivated by the recent success o

safetyarxiv-cs-ro
21 Apr 2026
Safety

Conformal Policy Control

DGX agent

arXiv:2603.02196v2 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm

safetyarxiv-cs-lg
17 Apr 2026
Safety

Beyond Conservative Automated Driving in Multi-Agent Scenarios via Coupled Model Predictive Control and Deep Reinforcement Learning

DGX agent

arXiv:2604.13891v1 Announce Type: new Abstract: Automated driving at unsignalized intersections is challenging due to complex multi-vehicle interactions and the need to balance safety and efficiency.

safetyarxiv-cs-ro
16 Apr 2026
Safety

4/5 We upgraded our original 3-line “be correct” prompt → a much more detailed prompt that enforced a hierarchy of constraints for correctne…

DGX agent

4/5 We upgraded our original 3-line “be correct” prompt → a much more detailed prompt that enforced a hierarchy of constraints for correctness, regression safety, and minimality. Basically, get the LL

safetyai21-labs--x
15 Apr 2026
Model Releases

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

DGX agent

arXiv:2512.20798v4 Announce Type: replace Abstract: As autonomous AI agents are deployed in high-stakes environments, ensuring their safety has become a paramount concern. Existing safety benchmarks p

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Cloud CISO Perspectives: How CISOs can pursue technical and cultural resilience (Q&A)

DGX agent

Welcome to the first Cloud CISO Perspectives for April 2026. Today, Thiébaut Meyer and Lia Wertheimer from Google Cloud’s Office of the CISO share Thiébaut’s conversation with Matt Rowe, chief securit

safetygoogle-cloud-ai
15 Apr 2026
Safety

Efficient Disruption of Criminal Networks through Multi-Objective Genetic Algorithms

DGX agent

arXiv:2604.09647v1 Announce Type: cross Abstract: Criminal networks, such as the Sicilian Mafia, pose substantial threats to public safety, national security, and economic stability. Outdated disrupti

safetyarxiv-cs-ai
14 Apr 2026
Safety

RoboStereo: Dual-Tower 4D Embodied World Models for Unified Policy Optimization

DGX agent

arXiv:2603.12639v2 Announce Type: replace Abstract: Scalable Embodied AI faces fundamental constraints due to prohibitive costs and safety risks of real-world interaction. While Embodied World Models

safetyarxiv-cs-cv
14 Apr 2026
Safety

Designing Safe and Accountable GenAI as a Learning Companion with Women Banned from Formal Education

DGX agent

arXiv:2604.07253v1 Announce Type: cross Abstract: In gender-restrictive and surveilled contexts, where access to formal education may be restricted for women, pursuing education involves safety and pr

safetyarxiv-cs-ai
10 Apr 2026
Safety

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

DGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

safetyarxiv-cs-lg
10 Apr 2026
Safety

Leveraging Wireless Sensor Networks for Real-Time Monitoring and Control of Industrial Environments

DGX agent

arXiv:2510.13820v3 Announce Type: replace-cross Abstract: This research proposes an extensive technique for monitoring and controlling the industrial parameters using Internet of Things (IoT) technolo

safetyarxiv-cs-ai
10 Apr 2026
Safety

Autonomous Driving with Priority-Ordered STL Specifications Under Multimodal Uncertainty

DGX agent

arXiv:2606.20336v2 Announce Type: replace Abstract: Autonomous vehicles must plan trajectories that satisfy multiple requirements, such as safety, traffic-rule compliance, and passenger comfort. Howev

safetyarxiv-cs-ro
11 Aug 2026
Safety

Pragmatic Attack Surface: Vulnerabilities of Implicit Context in Large Language Models

DGX agent

arXiv:2608.09551v1 Announce Type: new Abstract: In the era of large language models (LLMs), attackers often manipulate natural language to elicit unsafe or harmful outputs, creating a new natural lang

safetyarxiv-cs-cl
11 Aug 2026
Safety

SAFE-CHEM: Uncertainty-Aware Policy Switching for Robust Robotic Chemistry

DGX agent

arXiv:2608.09303v1 Announce Type: cross Abstract: The deployment of autonomous robotic systems in chemistry laboratories is accelerating experimental workflows and providing the foundational data for

safetyarxiv-cs-lg
11 Aug 2026
Model Releases

SafeSceneReason: A Multimodal Reasoning Benchmark Connecting Industrial Hazards with Accident Knowledge

DGX agent

arXiv:2608.09230v1 Announce Type: new Abstract: Industrial-safety understanding requires more than detecting workers, equipment, and personal protective equipment. Models must also assess compliance,

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding

DGX agent

arXiv:2512.09062v2 Announce Type: replace Abstract: Accurate 3D scene interpretation in active construction sites is essential for progress monitoring, safety assessment, and digital twin development.

safetyarxiv-cs-cv
11 Aug 2026
Safety

Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM

DGX agent

arXiv:2511.18721v4 Announce Type: replace-cross Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict 'k-unstable' assumption that

safetyarxiv-cs-ai
11 Aug 2026
Safety

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “…

DGX agent

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “who distributes it to the most people”: Instead of fighting

safetyyann-lecun--x
10 Aug 2026
Safety

Safe Evolution with Circuit Anchors

DGX agent

arXiv:2608.05158v1 Announce Type: new Abstract: In biological evolution, unconstrained mutation can lead to catastrophic outcomes: organisms may evolve enhanced capabilities while losing essential fun

safetyarxiv-cs-cl
7 Aug 2026
Safety

Optimal Constrained sc-LTL Planning in MDPs via Switching Policies

DGX agent

arXiv:2608.05021v1 Announce Type: new Abstract: We study the synthesis of optimal policies for planning problems on Markov decision processes with both objectives and safety constraints specified in c

safetyarxiv-cs-ro
6 Aug 2026
Safety

SafeLand: Safe Autonomous Landing in Unknown Environments with Bayesian Semantic Mapping

DGX agent

arXiv:2603.17430v2 Announce Type: replace Abstract: Autonomous landing of uncrewed aerial vehicles (UAVs) in unknown, dynamic environments poses significant safety challenges, particularly near people

safetyarxiv-cs-ro
6 Aug 2026
Safety

When Memory Lies: An Empirical Study of Spatial Memory Staleness in VLM Agents

DGX agent

arXiv:2608.04574v1 Announce Type: new Abstract: Memory-augmented VLM agents act on persistent spatial knowledge, yet that knowledge silently goes stale as the environment changes. We ask what happens

safetyarxiv-cs-cl
6 Aug 2026
Safety

Test-Time Scaling for Safe Text-Guided Image Generation via Intermediate Clean Estimates

DGX agent

arXiv:2608.03284v1 Announce Type: cross Abstract: Ensuring safety and policy compliance in text-to-image diffusion models remains a critical challenge, as benign or adversarial prompts can often elici

safetyarxiv-cs-ai
5 Aug 2026
Safety

Grounded Vision-Language Interpreter for Long-Horizon Bimanual Task and Motion Planning

DGX agent

arXiv:2506.03270v3 Announce Type: replace Abstract: While recent advances in vision-language models have accelerated language-guided robot planning, their black-box nature lacks the safety guarantees

safetyarxiv-cs-ro
4 Aug 2026
Safety

The model takes moderation policy as a plain-language question and returns a calibrated score. Text and images — one interface. Read the ful…

DGX agent

Mistral AI has released Shieldstral, a 3‑billion‑parameter open‑weight model for content safety that can run locally on device. It interprets plain‑language moderation queries and returns calibrated s

safetymistral-ai--x
4 Aug 2026
Safety

Beyond Component Testing: Validating Agentic AI Systems

DGX agent

arXiv:2607.29405v1 Announce Type: new Abstract: Agentic AI systems act through multi-step trajectories that combine planning, tool use, memory, interaction, and adaptation. This behavior stretches val

safetyarxiv-cs-ai
3 Aug 2026
Safety

RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving

DGX agent

arXiv:2602.07339v2 Announce Type: replace Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denoising process introduces a latency bottleneck fo

safetyarxiv-cs-ai
3 Aug 2026
Safety

TextCloak: Thwarting Unauthorized LLM Exploitation via RL-Driven Unlearnable Text

DGX agent

arXiv:2607.28862v1 Announce Type: cross Abstract: The rapid development of Large Language Models (LLMs) has led to significant advances across a wide range of language tasks, while simultaneously rais

safetyarxiv-cs-ai
3 Aug 2026
← Previous
1…1920212223…297
Next →