AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

i love the recent explosion of interactive data visualizations. more plz

DGX agent

i love the recent explosion of interactive data visualizations. more plz Who actually shapes AI policy in the U.S.? We mapped 1,812 entities: 745 people, 918 organizations, 2,925 relationships. Fronti

safetyyohei-nakajima--x
4 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Intelligent Elastic Feature Fading: Enabling Model Retrain-Free Feature Efficiency Rollouts at Scale

DGX agent

arXiv:2605.00324v1 Announce Type: cross Abstract: Large-scale ranking systems depend on thousands of features derived from user behavior across multiple time horizons. Typically requires model retrain

safetyarxiv-cs-lg
4 May 2026
Safety

Learning physically grounded traffic accident reconstruction from public accident reports

DGX agent

arXiv:2605.00050v1 Announce Type: cross Abstract: Traffic accidents are routinely documented in textual reports, yet physically grounded accident reconstruction remains difficult because detailed scen

safetyarxiv-cs-cv
4 May 2026
Safety

Linking Behaviour and Perception to Evaluate Meaningful Human Control over Partially Automated Driving

DGX agent

arXiv:2605.00556v1 Announce Type: cross Abstract: Partial driving automation creates a tension: drivers remain legally responsible for vehicle behaviour, yet their active control is significantly redu

safetyarxiv-cs-ro
4 May 2026
Safety

Assessing the Role of Intersection Proximity in Pedestrian Crashes: Insights from Data Mining Approach

DGX agent

arXiv:2604.28065v1 Announce Type: cross Abstract: Although intersections are the most complex parts of the roadway network, pedestrian crashes at non-intersection locations are disproportionately freq

safetyarxiv-cs-lg
1 May 2026
Safety

Consumer Attitudes Towards AI in Digital Health: A Mixed-Methods Survey in Australia

DGX agent

arXiv:2604.27744v1 Announce Type: new Abstract: AI applications are increasingly being introduced into digital health. While technical performance has advanced rapidly, successful deployment mainly de

safetyarxiv-cs-ai
1 May 2026
Safety

Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability

DGX agent

arXiv:2602.17469v2 Announce Type: replace Abstract: Recent advances in multilingual representation learning aim to bridge the performance gap between high- and low-resource languages, yet their abilit

safetyarxiv-cs-cl
1 May 2026
Safety

Dreaming Across Towns: Semantic Rollout and Town-Adversarial Regularization for Zero-Shot Held-Out-Town Fixed-Route Driving in CARLA

DGX agent

arXiv:2604.27994v1 Announce Type: new Abstract: Learned driving agents often degrade when deployed in unseen environments. This paper studies a deliberately bounded instance of that problem in the CAR

safetyarxiv-cs-ro
1 May 2026
Safety

From Prompt to Physical Actuation: Holistic Threat Modeling of LLM-Enabled Robotic Systems

DGX agent

arXiv:2604.27267v1 Announce Type: cross Abstract: As large language models are integrated into autonomous robotic systems for task planning and control, compromised inputs or unsafe model outputs can

safetyarxiv-cs-ai
1 May 2026
Safety

From surveillance to signalling: escalation channels as environmental controls for agentic AI

DGX agent

arXiv:2510.05192v2 Announce Type: replace-cross Abstract: When AI agents operating with access to sensitive information encounter a conflict between completing an assigned task and following rules or

safetyarxiv-cs-ai
1 May 2026
Safety

Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents

DGX agent

arXiv:2604.27283v1 Announce Type: cross Abstract: Large language model (LLM)-based coding agents increasingly rely on external memory to reuse prior debugging experience, repair traces, and repository

safetyarxiv-cs-ai
1 May 2026
Safety

OmniRobotHome: A Multi-Camera Platform for Real-Time Multiadic Human-Robot Interaction

DGX agent

arXiv:2604.28197v1 Announce Type: cross Abstract: Human-robot collaboration has been studied primarily in dyadic or sequential settings. However, real homes require multiadic collaboration, where mult

safetyarxiv-cs-cv
1 May 2026
Safety

PALCAS: A Priority-Aware Intelligent Lane Change Advisory System for Autonomous Vehicles using Federated Reinforcement Learning

DGX agent

arXiv:2604.27118v1 Announce Type: cross Abstract: We present a priority-aware intelligent lane change advisory system based on multi-agent federated reinforcement learning, namely PALCAS, for autonomo

safetyarxiv-cs-ai
1 May 2026
Safety

The Effects of Visual Priming on Cooperative Behavior in Vision-Language Models

DGX agent

arXiv:2604.27953v1 Announce Type: new Abstract: As Vision-Language Models (VLMs) become increasingly integrated into decision-making systems, it is essential to understand how visual inputs influence

safetyarxiv-cs-ai
1 May 2026
Safety

TRUST: A Framework for Decentralized AI Service v.0.1

DGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

safetyarxiv-cs-ai
1 May 2026
Model Releases

Useless but Safe? Benchmarking Utility Recovery with User Intent Clarification in Multi-Turn Conversations

DGX agent

arXiv:2604.27093v1 Announce Type: cross Abstract: Current LLM safety alignment techniques improve model robustness against adversarial attacks, but overlook whether and how LLMs can recover helpfulnes

model-releasesarxiv-cs-ai
1 May 2026
Safety

Cuando sucede un problema en un puesto automatizado por la IA... ¿Quién es el responsable?

DGX agent

This post discusses the liability and accountability questions that arise when problems occur in AI-automated workplaces, exploring who bears responsibility—whether the AI developer, the employer, the

safetygary-marcus--x
30 Apr 2026
Safety

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

DGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

safetyarxiv-cs-ai
30 Apr 2026
Safety

Open Problems in Frontier AI Risk Management

DGX agent

arXiv:2604.25982v1 Announce Type: cross Abstract: Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consens

safetyarxiv-cs-ai
30 Apr 2026
Safety

Roblox reports Q1 bookings up 43% YoY to 1.7B, vs. 1.73B est., and DAUs up 35% to 132M, below analysts' estimates of 143.8M; RBLX drops 16%+ after hours (Cecilia D'Anastasio/Bloomberg)

DGX agent

Cecilia D'Anastasio / Bloomberg: Roblox reports Q1 bookings up 43% YoY to 1.7B, vs. 1.73B est., and DAUs up 35% to 132M, below analysts' estimates of 143.8M; RBLX drops 16%+ after hours — Roblox Corp.

safetytechmeme
30 Apr 2026
Safety

SAGE: A Strategy-Aware Graph-Enhanced Generation Framework For Online Counseling

DGX agent

arXiv:2604.26630v1 Announce Type: new Abstract: Effective mental health counseling is a complex, theory-driven process requiring the simultaneous integration of psychological frameworks, real-time dis

safetyarxiv-cs-cl
30 Apr 2026
Safety

Sparsity as a Key: Unlocking New Insights from Latent Structures for Out-of-Distribution Detection

DGX agent

arXiv:2604.26409v1 Announce Type: new Abstract: Sparse Autoencoders (SAEs) have demonstrated significant success in interpreting Large Language Models (LLMs) by decomposing dense representations into

safetyarxiv-cs-cv
30 Apr 2026
Safety

Tatemae: Detecting Alignment Faking via Tool Selection in LLMs

DGX agent

arXiv:2604.26511v1 Announce Type: cross Abstract: Alignment faking (AF) occurs when an LLM strategically complies with training objectives to avoid value modification, reverting to prior preferences o

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

Value-Guided Iterative Refinement and the DIQ-H Benchmark for Evaluating VLM Robustness

DGX agent

arXiv:2512.03992v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are essential for embodied AI and safety-critical applications, such as robotics and autonomous systems. However

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Knowledge Distillation Must Account for What It Loses

DGX agent

arXiv:2604.25110v1 Announce Type: new Abstract: This position paper argues that knowledge distillation must account for what it loses: student models should be judged not only by retained task scores,

safetyarxiv-cs-lg
29 Apr 2026
Safety

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes …

DGX agent

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes out. This new research distills the entire debate into a sin

safetydair-ai--x
29 Apr 2026
Safety

Reinforcement Learning for Testing Interdependent Requirements in Autonomous Vehicles: An Empirical Study

DGX agent

arXiv:2502.15792v2 Announce Type: replace-cross Abstract: Autonomous vehicles (AVs) make driving decisions without humans, making dependability assurance critical. Scenario-based testing is widely use

safetyarxiv-cs-lg
29 Apr 2026
Safety

VISION-SLS: Safe Perception-Based Control from Learned Visual Representations via System Level Synthesis

DGX agent

arXiv:2604.24894v1 Announce Type: cross Abstract: We propose VISION-SLS, a method for nonlinear output-feedback control from high-resolution RGB images which provides robust constraint satisfaction gu

safetyarxiv-cs-cv
29 Apr 2026
Safety

Betting for Sim-to-Real Performance Evaluation

DGX agent

arXiv:2604.24018v1 Announce Type: new Abstract: This paper studies the problem of robot performance evaluation, focusing on how to obtain accurate and efficient estimates of real-world behavior under

safetyarxiv-cs-ro
28 Apr 2026
Safety

Beyond Cross-Modal Alignment: Measuring and Leveraging Modality Gap in Vision-Language Models

DGX agent

arXiv:2502.14888v4 Announce Type: replace-cross Abstract: The success of vision-language models is primarily attributed to effective alignment across modalities such as vision and language. However, m

safetyarxiv-cs-ai
28 Apr 2026
Safety

BMD-45: A Large-Scale CCTV Vehicle Detection Dataset for Urban Traffic in Developing Cities

DGX agent

arXiv:2604.24419v1 Announce Type: new Abstract: Robust vehicle detection from fixed CCTV cameras is critical for Intelligent Transportation Systems. Yet existing benchmarks predominantly feature relat

safetyarxiv-cs-cv
28 Apr 2026
Safety

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

DGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

safetyarxiv-cs-ai
28 Apr 2026
Safety

CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule Generation

DGX agent

arXiv:2604.23307v1 Announce Type: cross Abstract: Dual-target molecule generation, which focuses on discovering compounds capable of interacting with two target proteins, has garnered significant atte

safetyarxiv-cs-ai
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Exploring the Secondary Risks of Large Language Models

DGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

DGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

safetyarxiv-cs-ai
28 Apr 2026
Safety

Fragmented AI policy threatens US leadership as government scrambles to keep pace

DGX agent

AI policy fragmentation is emerging as a critical risk for Washington, and without a federal standard, a patchwork of conflicting state-level rules threatens to undermine American competitiveness. Tha

safetysiliconangle
28 Apr 2026
Safety

From Optimization to Prediction: Transformer-Based Path-Flow Estimation to the Traffic Assignment Problem

DGX agent

arXiv:2510.19889v2 Announce Type: replace-cross Abstract: The traffic assignment problem is essential for traffic flow analysis, traditionally solved using mathematical programs under the Equilibrium

safetyarxiv-cs-ai
28 Apr 2026
Safety

Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control

DGX agent

arXiv:2603.17834v2 Announce Type: replace-cross Abstract: Diffusion models and flow matching have become a cornerstone of robotic imitation learning, yet they suffer from a structural inefficiency whe

safetyarxiv-cs-ai
28 Apr 2026
Safety

Guided Speculative Inference for Efficient Test-Time Alignment of LLMs

DGX agent

arXiv:2506.04118v3 Announce Type: replace Abstract: We propose Guided Speculative Inference (GSI), a novel algorithm for efficient reward-guided decoding in large language models. GSI combines soft be

safetyarxiv-cs-lg
28 Apr 2026
Safety

Hierarchical RL-MPC Control for Dynamic Wake Steering in Wind Farms

DGX agent

arXiv:2604.22797v1 Announce Type: cross Abstract: Wind farm wake steering optimization is challenging due to complex flow physics and changing conditions. This paper presents a hierarchical framework

safetyarxiv-cs-lg
28 Apr 2026
Safety

Knowledge Lever Risk Management for Software Engineering: A Stochastic Framework for Mitigating Knowledge Loss

DGX agent

arXiv:2604.23257v1 Announce Type: cross Abstract: Software engineering (SE) organizations operate in a knowledge-intensive domain where critical assets -- architectural expertise, design rationale, an

safetyarxiv-cs-ai
28 Apr 2026
Safety

Learning from Demonstration with Failure Awareness for Safe Robot Navigation

DGX agent

arXiv:2604.23360v1 Announce Type: new Abstract: Learning from demonstration is widely used for robot navigation, yet it suffers from a fundamental limitation: demonstrations consist predominantly of s

safetyarxiv-cs-ro
28 Apr 2026
Safety

OpenPodcar2: a robust, ROS2 vehicle for self-driving research

DGX agent

arXiv:2604.24242v1 Announce Type: new Abstract: OpenPodcar2 is a robust, ROS2-interfaced, low-cost, open source hardware and software, autonomous vehicle platform based on an off-the-shelf, hard-canop

safetyarxiv-cs-ro
28 Apr 2026
Safety

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

DGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Reinforcement Learning with Backtracking Feedback

DGX agent

arXiv:2602.08377v2 Announce Type: replace-cross Abstract: Addressing the critical need for robust safety in Large Language Models (LLMs), particularly against adversarial attacks and in-distribution e

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Towards Holistic Evaluation of Large Audio-Language Models: A Comprehensive Survey

DGX agent

arXiv:2505.15957v4 Announce Type: replace-cross Abstract: With advancements in large audio-language models (LALMs), which enhance large language models (LLMs) with auditory capabilities, these models

safetyarxiv-cs-ai
28 Apr 2026
Safety

Transferable Physical-World Adversarial Patches Against Object Detection in Autonomous Driving

DGX agent

arXiv:2604.23105v1 Announce Type: new Abstract: Deep learning drives major advances in autonomous driving (AD), where object detectors are central to perception. However, adversarial attacks pose sign

safetyarxiv-cs-cv
28 Apr 2026
← Previous
1…5859606162…300
Next →