AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Linking Behaviour and Perception to Evaluate Meaningful Human Control over Partially Automated Driving

DGX agent

arXiv:2605.00556v1 Announce Type: cross Abstract: Partial driving automation creates a tension: drivers remain legally responsible for vehicle behaviour, yet their active control is significantly redu

safetyarxiv-cs-ro
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Assessing the Role of Intersection Proximity in Pedestrian Crashes: Insights from Data Mining Approach

DGX agent

arXiv:2604.28065v1 Announce Type: cross Abstract: Although intersections are the most complex parts of the roadway network, pedestrian crashes at non-intersection locations are disproportionately freq

safetyarxiv-cs-lg
1 May 2026
Safety

Consumer Attitudes Towards AI in Digital Health: A Mixed-Methods Survey in Australia

DGX agent

arXiv:2604.27744v1 Announce Type: new Abstract: AI applications are increasingly being introduced into digital health. While technical performance has advanced rapidly, successful deployment mainly de

safetyarxiv-cs-ai
1 May 2026
Safety

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability

DGX agent

arXiv:2602.17469v2 Announce Type: replace Abstract: Recent advances in multilingual representation learning aim to bridge the performance gap between high- and low-resource languages, yet their abilit

safetyarxiv-cs-cl
1 May 2026
Safety

Dreaming Across Towns: Semantic Rollout and Town-Adversarial Regularization for Zero-Shot Held-Out-Town Fixed-Route Driving in CARLA

DGX agent

arXiv:2604.27994v1 Announce Type: new Abstract: Learned driving agents often degrade when deployed in unseen environments. This paper studies a deliberately bounded instance of that problem in the CAR

safetyarxiv-cs-ro
1 May 2026
Safety

From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

DGX agent

arXiv:2604.27267v1 Announce Type: cross Abstract: As large language models are integrated into autonomous robotic systems for task planning and control, compromised inputs or unsafe model outputs can

safetyarxiv-cs-ai
1 May 2026
Safety

From surveillance to signalling: escalation channels as environmental controls for agentic AI

DGX agent

arXiv:2510.05192v2 Announce Type: replace-cross Abstract: When AI agents operating with access to sensitive information encounter a conflict between completing an assigned task and following rules or

safetyarxiv-cs-ai
1 May 2026
Safety

Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents

DGX agent

arXiv:2604.27283v1 Announce Type: cross Abstract: Large language model (LLM)-based coding agents increasingly rely on external memory to reuse prior debugging experience, repair traces, and repository

safetyarxiv-cs-ai
1 May 2026
Safety

OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction

DGX agent

arXiv:2604.28197v1 Announce Type: cross Abstract: Human-robot collaboration has been studied primarily in dyadic or sequential settings. However, real homes require multiadic collaboration, where mult

safetyarxiv-cs-cv
1 May 2026
Safety

PALCAS: A Priority-Aware Intelligent Lane Change Advisory System for Autonomous Vehicles using Federated Reinforcement Learning

DGX agent

arXiv:2604.27118v1 Announce Type: cross Abstract: We present a priority-aware intelligent lane change advisory system based on multi-agent federated reinforcement learning, namely PALCAS, for autonomo

safetyarxiv-cs-ai
1 May 2026
Safety

The Effects of Visual Priming on Cooperative Behavior in Vision-Language Models

DGX agent

arXiv:2604.27953v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become increasingly integrated into decision-making systems, it is essential to understand how visual inputs influence

safetyarxiv-cs-ai
1 May 2026
Safety

TRUST: A Framework for Decentralized AI Service v.0.1

DGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

safetyarxiv-cs-ai
1 May 2026
Model Releases

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

DGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

model-releasesarxiv-cs-ai
1 May 2026
Safety

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

DGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

safetyarxiv-cs-ai
30 Apr 2026
Safety

Open Problems in Frontier AI Risk Management

DGX agent

arXiv:2604.25982v1 Announce Type: cross Abstract: Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consens

safetyarxiv-cs-ai
30 Apr 2026
Safety

SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling

DGX agent

arXiv:2604.26630v1 Announce Type: new Abstract: Effective mental health counseling is a complex, theory-driven process requiring the simultaneous integration of psychological frameworks, real-time dis

safetyarxiv-cs-cl
30 Apr 2026
Safety

Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection

DGX agent

arXiv:2604.26409v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have demonstrated significant success in interpreting Large Language Models (LLMs) by decomposing dense representations into

safetyarxiv-cs-cv
30 Apr 2026
Safety

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

DGX agent

arXiv:2604.26511v1 Announce Type: cross Abstract: Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences o

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

DGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Knowledge Distillation Must Account for What It Loses

DGX agent

arXiv:2604.25110v1 Announce Type: new Abstract: This position paper argues that knowledge distillation must account for what it loses: student models should be judged not only by retained task scores,

safetyarxiv-cs-lg
29 Apr 2026
Safety

Reinforcement Learning for Testing Interdependent Requirements in Autonomous Vehicles: An Empirical Study

DGX agent

arXiv:2502.15792v2 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) make driving decisions without humans, making dependability assurance critical. Scenario-based testing is widely use

safetyarxiv-cs-lg
29 Apr 2026
Safety

VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis

DGX agent

arXiv:2604.24894v1 Announce Type: cross Abstract: We propose VISION-SLS, a method for nonlinear output-feedback control from high-resolution RGB images which provides robust constraint satisfaction gu

safetyarxiv-cs-cv
29 Apr 2026
Safety

Betting for Sim-to-Real Performance Evaluation

DGX agent

arXiv:2604.24018v1 Announce Type: new Abstract: This paper studies the problem of robot performance evaluation, focusing on how to obtain accurate and efficient estimates of real-world behavior under

safetyarxiv-cs-ro
28 Apr 2026
Safety

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models

DGX agent

arXiv:2502.14888v4 Announce Type: replace-cross Abstract: The success of vision-language models is primarily attributed to effective alignment across modalities such as vision and language. However, m

safetyarxiv-cs-ai
28 Apr 2026
Safety

BMD-45: A Large-Scale CCTV Vehicle Detection Dataset for Urban Traffic in Developing Cities

DGX agent

arXiv:2604.24419v1 Announce Type: new Abstract: Robust vehicle detection from fixed CCTV cameras is critical for Intelligent Transportation Systems. Yet existing benchmarks predominantly feature relat

safetyarxiv-cs-cv
28 Apr 2026
Safety

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

DGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

safetyarxiv-cs-ai
28 Apr 2026
Safety

CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule Generation

DGX agent

arXiv:2604.23307v1 Announce Type: cross Abstract: Dual-target molecule generation, which focuses on discovering compounds capable of interacting with two target proteins, has garnered significant atte

safetyarxiv-cs-ai
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Exploring the Secondary Risks of Large Language Models

DGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

DGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

safetyarxiv-cs-ai
28 Apr 2026
Safety

From Optimization to Prediction: Transformer-Based Path-Flow Estimation to the Traffic Assignment Problem

DGX agent

arXiv:2510.19889v2 Announce Type: replace-cross Abstract: The traffic assignment problem is essential for traffic flow analysis, traditionally solved using mathematical programs under the Equilibrium

safetyarxiv-cs-ai
28 Apr 2026
Safety

Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control

DGX agent

arXiv:2603.17834v2 Announce Type: replace-cross Abstract: Diffusion models and flow matching have become a cornerstone of robotic imitation learning, yet they suffer from a structural inefficiency whe

safetyarxiv-cs-ai
28 Apr 2026
Safety

Guided Speculative Inference for Efficient Test-Time Alignment of LLMs

DGX agent

arXiv:2506.04118v3 Announce Type: replace Abstract: We propose Guided Speculative Inference (GSI), a novel algorithm for efficient reward-guided decoding in large language models. GSI combines soft be

safetyarxiv-cs-lg
28 Apr 2026
Safety

Hierarchical RL-MPC Control for Dynamic Wake Steering in Wind Farms

DGX agent

arXiv:2604.22797v1 Announce Type: cross Abstract: Wind farm wake steering optimization is challenging due to complex flow physics and changing conditions. This paper presents a hierarchical framework

safetyarxiv-cs-lg
28 Apr 2026
Safety

Knowledge Lever Risk Management for Software Engineering: A Stochastic Framework for Mitigating Knowledge Loss

DGX agent

arXiv:2604.23257v1 Announce Type: cross Abstract: Software engineering (SE) organizations operate in a knowledge-intensive domain where critical assets -- architectural expertise, design rationale, an

safetyarxiv-cs-ai
28 Apr 2026
Safety

Learning from Demonstration with Failure Awareness for Safe Robot Navigation

DGX agent

arXiv:2604.23360v1 Announce Type: new Abstract: Learning from demonstration is widely used for robot navigation, yet it suffers from a fundamental limitation: demonstrations consist predominantly of s

safetyarxiv-cs-ro
28 Apr 2026
Safety

OpenPodcar2: a robust, ROS2 vehicle for self-driving research

DGX agent

arXiv:2604.24242v1 Announce Type: new Abstract: OpenPodcar2 is a robust, ROS2-interfaced, low-cost, open source hardware and software, autonomous vehicle platform based on an off-the-shelf, hard-canop

safetyarxiv-cs-ro
28 Apr 2026
Safety

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

DGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Reinforcement Learning with Backtracking Feedback

DGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

DGX agent

arXiv:2505.15957v4 Announce Type: replace-cross Abstract: With advancements in large audio-language models (LALMs), which enhance large language models (LLMs) with auditory capabilities, these models

safetyarxiv-cs-ai
28 Apr 2026
Safety

Transferable Physical-World Adversarial Patches Against Object Detection in Autonomous Driving

DGX agent

arXiv:2604.23105v1 Announce Type: new Abstract: Deep learning drives major advances in autonomous driving (AD), where object detectors are central to perception. However, adversarial attacks pose sign

safetyarxiv-cs-cv
28 Apr 2026
Safety

UGAF-ITS: A Standards Harmonization Framework and Validation Tool for Multi-Framework AI Governance in Distributed Intelligent Transportation Systems

DGX agent

arXiv:2604.22789v1 Announce Type: cross Abstract: Organizations deploying AI-enabled Intelligent Transportation Systems face fragmented governance: ISO/IEC 42001 demands a certifiable management syste

safetyarxiv-cs-ai
28 Apr 2026
Safety

UniAda: Universal Adaptive Multi-objective Adversarial Attack for End-to-End Autonomous Driving Systems

DGX agent

arXiv:2604.23362v1 Announce Type: cross Abstract: Adversarial attacks play a pivotal role in testing and improving the reliability of deep learning (DL) systems. Existing literature has demonstrated t

safetyarxiv-cs-lg
28 Apr 2026
Safety

Verifying Quantized GNNs With Readout Is Decidable But Highly Intractable

DGX agent

arXiv:2510.08045v2 Announce Type: replace-cross Abstract: We introduce a logical language for reasoning about quantized aggregate-combine graph neural networks with global readout (ACR-GNNs). We provi

safetyarxiv-cs-ai
28 Apr 2026
Safety

When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning

DGX agent

arXiv:2604.22873v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) can learn effective policies from fixed datasets, but deployment objectives may change after training, and in many

safetyarxiv-cs-ai
28 Apr 2026
Safety

A Co-Evolutionary Theory of Human-AI Coexistence: Mutualism, Governance, and Dynamics in Complex Societies

DGX agent

arXiv:2604.22227v1 Announce Type: cross Abstract: Classical robot ethics is often framed around obedience, most famously through Asimov's laws. This framing is too narrow for contemporary AI systems,

safetyarxiv-cs-ai
27 Apr 2026
Safety

How Large Language Models Balance Internal Knowledge with User and Document Assertions

DGX agent

arXiv:2604.22193v1 Announce Type: new Abstract: Large language models (LLMs) often need to balance their internal parametric knowledge with external information, such as user beliefs and content from

safetyarxiv-cs-cl
27 Apr 2026
Safety

Learning-augmented robotic automation for real-world manufacturing

DGX agent

arXiv:2604.22235v1 Announce Type: cross Abstract: Industrial robots are widely used in manufacturing, yet most manipulation still depends on fixed waypoint scripts that are brittle to environmental ch

safetyarxiv-cs-ai
27 Apr 2026
← Previous
1…5253545556…257
Next →