AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
5 May 2026

PRCD-MAP: Learning How Much to Trust Imperfect Priors in Causal Discovery

SafetyDGX agent

arXiv:2605.01669v1 Announce Type: cross Abstract: External priors of unknown reliability create a brittle trade-off in causal discovery: blind trust amplifies errors, blind rejection wastes signal. Re

Protein-Conditioned Multi-Objective Reinforcement Learning for Full-Length mRNA Design

SafetyDGX agent

arXiv:2605.01513v1 Announce Type: new Abstract: Designing therapeutic messenger RNA (mRNA) requires creating full-length transcripts that carefully balance stability, translation efficiency, and immun

Reinforcement Learning from Compiler and Language Server Feedback

SafetyDGX agent

arXiv:2510.22907v2 Announce Type: replace Abstract: Coding agents fail when text-level guesses outrun program facts: they hallucinate APIs, drift to the wrong symbol, and apply edits without evidence

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Semantic Risk-Aware Heuristic Planning for Robotic Navigation in Dynamic Environments: An LLM-Inspired Approach

SafetyDGX agent

arXiv:2605.02862v1 Announce Type: new Abstract: The integration of Large Language Model (LLM) reasoning principles into classical robot path planning represents a rapidly emerging research direction.

SurgTEMP: Temporal-Aware Surgical Video Question Answering with Text-guided Visual Memory for Laparoscopic Cholecystectomy

SafetyDGX agent

arXiv:2603.29962v3 Announce Type: replace Abstract: Surgical procedures are inherently complex and risky, requiring extensive expertise and constant focus to navigate evolving intraoperative scenes. C

The crux of the Musk-OpenAI trial so far is that Brockman keeps acting as if selling chatbots and API access for profit (which is what they …

SafetyDGX agent

The crux of the Musk-OpenAI trial so far is that Brockman keeps acting as if selling chatbots and API access for profit (which is what they do now) is exactly the same as the original mission (which w

Thermal Imaging for Contactless Cardiorespiratory and Sudomotor Response Monitoring

SafetyDGX agent

arXiv:2602.12361v2 Announce Type: replace Abstract: Human-machine interfaces in industrial automation need sensing modules that monitor operator actions and physiological state. This is important in f

Training-Free Adaptive 360-degree Video Streaming via Semantic Potential Fields

SafetyDGX agent

arXiv:2603.20999v2 Announce Type: replace-cross Abstract: Adaptive 360{eg} video streaming for teleoperation faces two coupled challenges: viewport prediction under uncertain gaze patterns and bitrate

Virtual Speech Therapist: A Clinician-in-the-Loop AI Speech Therapy Agent for Personalized and Supervised Therapy

SafetyDGX agent

arXiv:2605.01101v1 Announce Type: cross Abstract: This paper develops Virtual Speech Therapist (VST), an intelligent agent-based platform that streamlines stuttering assessment and delivers customized

4 May 2026

Conformalized Quantum DeepONet Ensembles for Scalable Operator Learning with Distribution-Free Uncertainty

SafetyDGX agent

arXiv:2605.00330v1 Announce Type: new Abstract: Operator learning enables fast surrogate modeling of high-dimensional dynamical systems, but existing approaches face two fundamental limitations: quadr

Exploring LLM biases to manipulate AI search overview

SafetyDGX agent

arXiv:2605.00012v1 Announce Type: cross Abstract: Modern large language models (LLMs) are used in many business applications in general, and specifically in web search systems and applications that ge

i love the recent explosion of interactive data visualizations. more plz

SafetyDGX agent

i love the recent explosion of interactive data visualizations. more plz Who actually shapes AI policy in the U.S.? We mapped 1,812 entities: 745 people, 918 organizations, 2,925 relationships. Fronti

Intelligent Elastic Feature Fading: Enabling Model Retrain-Free Feature Efficiency Rollouts at Scale

SafetyDGX agent

arXiv:2605.00324v1 Announce Type: cross Abstract: Large-scale ranking systems depend on thousands of features derived from user behavior across multiple time horizons. Typically requires model retrain

Learning physically grounded traffic accident reconstruction from public accident reports

SafetyDGX agent

arXiv:2605.00050v1 Announce Type: cross Abstract: Traffic accidents are routinely documented in textual reports, yet physically grounded accident reconstruction remains difficult because detailed scen

Linking Behaviour and Perception to Evaluate Meaningful Human Control over Partially Automated Driving

SafetyDGX agent

arXiv:2605.00556v1 Announce Type: cross Abstract: Partial driving automation creates a tension: drivers remain legally responsible for vehicle behaviour, yet their active control is significantly redu

1 May 2026

Assessing the Role of Intersection Proximity in Pedestrian Crashes: Insights from Data Mining Approach

SafetyDGX agent

arXiv:2604.28065v1 Announce Type: cross Abstract: Although intersections are the most complex parts of the roadway network, pedestrian crashes at non-intersection locations are disproportionately freq

Consumer Attitudes Towards AI in Digital Health: A Mixed-Methods Survey in Australia

SafetyDGX agent

arXiv:2604.27744v1 Announce Type: new Abstract: AI applications are increasingly being introduced into digital health. While technical performance has advanced rapidly, successful deployment mainly de

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability

SafetyDGX agent

arXiv:2602.17469v2 Announce Type: replace Abstract: Recent advances in multilingual representation learning aim to bridge the performance gap between high- and low-resource languages, yet their abilit

Dreaming Across Towns: Semantic Rollout and Town-Adversarial Regularization for Zero-Shot Held-Out-Town Fixed-Route Driving in CARLA

SafetyDGX agent

arXiv:2604.27994v1 Announce Type: new Abstract: Learned driving agents often degrade when deployed in unseen environments. This paper studies a deliberately bounded instance of that problem in the CAR

From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

SafetyDGX agent

arXiv:2604.27267v1 Announce Type: cross Abstract: As large language models are integrated into autonomous robotic systems for task planning and control, compromised inputs or unsafe model outputs can

From surveillance to signalling: escalation channels as environmental controls for agentic AI

SafetyDGX agent

arXiv:2510.05192v2 Announce Type: replace-cross Abstract: When AI agents operating with access to sensitive information encounter a conflict between completing an assigned task and following rules or

Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents

SafetyDGX agent

arXiv:2604.27283v1 Announce Type: cross Abstract: Large language model (LLM)-based coding agents increasingly rely on external memory to reuse prior debugging experience, repair traces, and repository

OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction

SafetyDGX agent

arXiv:2604.28197v1 Announce Type: cross Abstract: Human-robot collaboration has been studied primarily in dyadic or sequential settings. However, real homes require multiadic collaboration, where mult

PALCAS: A Priority-Aware Intelligent Lane Change Advisory System for Autonomous Vehicles using Federated Reinforcement Learning

SafetyDGX agent

arXiv:2604.27118v1 Announce Type: cross Abstract: We present a priority-aware intelligent lane change advisory system based on multi-agent federated reinforcement learning, namely PALCAS, for autonomo

The Effects of Visual Priming on Cooperative Behavior in Vision-Language Models

SafetyDGX agent

arXiv:2604.27953v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become increasingly integrated into decision-making systems, it is essential to understand how visual inputs influence

TRUST: A Framework for Decentralized AI Service v.0.1

SafetyDGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

Model ReleasesDGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

30 Apr 2026

Cuando sucede un problema en un puesto automatizado por la IA... ¿Quién es el responsable?

SafetyDGX agent

This post discusses the liability and accountability questions that arise when problems occur in AI-automated workplaces, exploring who bears responsibility—whether the AI developer, the employer, the

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

SafetyDGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

Open Problems in Frontier AI Risk Management

SafetyDGX agent

arXiv:2604.25982v1 Announce Type: cross Abstract: Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consens

Roblox reports Q1 bookings up 43% YoY to 1.7B, vs. 1.73B est., and DAUs up 35% to 132M, below analysts' estimates of 143.8M; RBLX drops 16%+ after hours (Cecilia D'Anastasio/Bloomberg)

SafetyDGX agent

Cecilia D'Anastasio / Bloomberg: Roblox reports Q1 bookings up 43% YoY to 1.7B, vs. 1.73B est., and DAUs up 35% to 132M, below analysts' estimates of 143.8M; RBLX drops 16%+ after hours — Roblox Corp.

SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling

SafetyDGX agent

arXiv:2604.26630v1 Announce Type: new Abstract: Effective mental health counseling is a complex, theory-driven process requiring the simultaneous integration of psychological frameworks, real-time dis

Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection

SafetyDGX agent

arXiv:2604.26409v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have demonstrated significant success in interpreting Large Language Models (LLMs) by decomposing dense representations into

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

SafetyDGX agent

arXiv:2604.26511v1 Announce Type: cross Abstract: Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences o

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

Model ReleasesDGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

29 Apr 2026

Knowledge Distillation Must Account for What It Loses

SafetyDGX agent

arXiv:2604.25110v1 Announce Type: new Abstract: This position paper argues that knowledge distillation must account for what it loses: student models should be judged not only by retained task scores,

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes …

SafetyDGX agent

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes out. This new research distills the entire debate into a sin

Reinforcement Learning for Testing Interdependent Requirements in Autonomous Vehicles: An Empirical Study

SafetyDGX agent

arXiv:2502.15792v2 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) make driving decisions without humans, making dependability assurance critical. Scenario-based testing is widely use

VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis

SafetyDGX agent

arXiv:2604.24894v1 Announce Type: cross Abstract: We propose VISION-SLS, a method for nonlinear output-feedback control from high-resolution RGB images which provides robust constraint satisfaction gu

28 Apr 2026

Betting for Sim-to-Real Performance Evaluation

SafetyDGX agent

arXiv:2604.24018v1 Announce Type: new Abstract: This paper studies the problem of robot performance evaluation, focusing on how to obtain accurate and efficient estimates of real-world behavior under

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models

SafetyDGX agent

arXiv:2502.14888v4 Announce Type: replace-cross Abstract: The success of vision-language models is primarily attributed to effective alignment across modalities such as vision and language. However, m

BMD-45: A Large-Scale CCTV Vehicle Detection Dataset for Urban Traffic in Developing Cities

SafetyDGX agent

arXiv:2604.24419v1 Announce Type: new Abstract: Robust vehicle detection from fixed CCTV cameras is critical for Intelligent Transportation Systems. Yet existing benchmarks predominantly feature relat

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

SafetyDGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule Generation

SafetyDGX agent

arXiv:2604.23307v1 Announce Type: cross Abstract: Dual-target molecule generation, which focuses on discovering compounds capable of interacting with two target proteins, has garnered significant atte

Discovering Failure Modes in Vision-Language Models using RL

SafetyDGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

Exploring the Secondary Risks of Large Language Models

Model ReleasesDGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

SafetyDGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

Fragmented AI policy threatens US leadership as government scrambles to keep pace

SafetyDGX agent

AI policy fragmentation is emerging as a critical risk for Washington, and without a federal standard, a patchwork of conflicting state-level rules threatens to undermine American competitiveness. Tha

From Optimization to Prediction: Transformer-Based Path-Flow Estimation to the Traffic Assignment Problem

SafetyDGX agent

arXiv:2510.19889v2 Announce Type: replace-cross Abstract: The traffic assignment problem is essential for traffic flow analysis, traditionally solved using mathematical programs under the Equilibrium

Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control

SafetyDGX agent

arXiv:2603.17834v2 Announce Type: replace-cross Abstract: Diffusion models and flow matching have become a cornerstone of robotic imitation learning, yet they suffer from a structural inefficiency whe

Guided Speculative Inference for Efficient Test-Time Alignment of LLMs

SafetyDGX agent

arXiv:2506.04118v3 Announce Type: replace Abstract: We propose Guided Speculative Inference (GSI), a novel algorithm for efficient reward-guided decoding in large language models. GSI combines soft be

Hierarchical RL-MPC Control for Dynamic Wake Steering in Wind Farms

SafetyDGX agent

arXiv:2604.22797v1 Announce Type: cross Abstract: Wind farm wake steering optimization is challenging due to complex flow physics and changing conditions. This paper presents a hierarchical framework

Knowledge Lever Risk Management for Software Engineering: A Stochastic Framework for Mitigating Knowledge Loss

SafetyDGX agent

arXiv:2604.23257v1 Announce Type: cross Abstract: Software engineering (SE) organizations operate in a knowledge-intensive domain where critical assets -- architectural expertise, design rationale, an

Learning from Demonstration with Failure Awareness for Safe Robot Navigation

SafetyDGX agent

arXiv:2604.23360v1 Announce Type: new Abstract: Learning from demonstration is widely used for robot navigation, yet it suffers from a fundamental limitation: demonstrations consist predominantly of s

OpenPodcar2: a robust, ROS2 vehicle for self-driving research

SafetyDGX agent

arXiv:2604.24242v1 Announce Type: new Abstract: OpenPodcar2 is a robust, ROS2-interfaced, low-cost, open source hardware and software, autonomous vehicle platform based on an off-the-shelf, hard-canop

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

SafetyDGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

Reinforcement Learning with Backtracking Feedback

Model ReleasesDGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

SafetyDGX agent

arXiv:2505.15957v4 Announce Type: replace-cross Abstract: With advancements in large audio-language models (LALMs), which enhance large language models (LLMs) with auditory capabilities, these models

Transferable Physical-World Adversarial Patches Against Object Detection in Autonomous Driving

SafetyDGX agent

arXiv:2604.23105v1 Announce Type: new Abstract: Deep learning drives major advances in autonomous driving (AD), where object detectors are central to perception. However, adversarial attacks pose sign

UGAF-ITS: A Standards Harmonization Framework and Validation Tool for Multi-Framework AI Governance in Distributed Intelligent Transportation Systems

SafetyDGX agent

arXiv:2604.22789v1 Announce Type: cross Abstract: Organizations deploying AI-enabled Intelligent Transportation Systems face fragmented governance: ISO/IEC 42001 demands a certifiable management syste

← Previous
1…4647484950…240
Next →