AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
Safety

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

DGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

safetyarxiv-cs-ai
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

For a long time, academic researchers being at the cutting edge of new technologies has been a great social equilibrium. Neutral, unbiased t…

DGX agent

For a long time, academic researchers being at the cutting edge of new technologies has been a great social equilibrium. Neutral, unbiased technologists have been the people to spread new ideas to the

safetyclem-delangue--x
19 May 2026
Safety

Guided Reinforcement Learning for Omnidirectional 3D Jumping in Quadruped Robots

DGX agent

arXiv:2507.16481v3 Announce Type: replace Abstract: Jumping poses a significant challenge for quadruped robots, despite being crucial for many operational scenarios. While optimisation methods exist f

safetyarxiv-cs-ro
19 May 2026
Safety

Helping Customers in Distress: An LLM-powered Agent that Converses, Probes, and Routes

DGX agent

arXiv:2605.16268v1 Announce Type: cross Abstract: Banks receive millions of reports of fraud, scams, and disputed transactions every year, making it challenging to accurately direct customers to the a

safetyarxiv-cs-ai
19 May 2026
Safety

Learning-Based Adaptive Control for Surgical Robotic Exposure Task on Deformable Tissues

DGX agent

arXiv:2605.17927v1 Announce Type: new Abstract: In various surgical procedures, regions of interest (ROIs) such as organs or lesions are often occluded by overlying tissues, requiring surgeons to achi

safetyarxiv-cs-ro
19 May 2026
Safety

On Safer Reinforcement Learning for Sedation and Analgesia in Intensive Care

DGX agent

arXiv:2601.23154v2 Announce Type: replace-cross Abstract: Pain management in intensive care usually involves complex trade-offs, since both inadequate and excessive treatment can compromise patient sa

safetyarxiv-cs-ai
19 May 2026
Safety

Online Learnability of Chain-of-Thought Verifiers: Soundness and Completeness Trade-offs

DGX agent

arXiv:2603.03538v3 Announce Type: replace Abstract: Large Language Models (LLMs) with chain-of-thought generation have demonstrated great potential for solving complex reasoning and planning tasks. Ho

safetyarxiv-cs-lg
19 May 2026
Safety

PQR: A Framework to Generate Diverse and Realistic User Queries that Elicit QA Agent Failures

DGX agent

arXiv:2605.16551v1 Announce Type: new Abstract: Evaluating LLM-based agents remains challenging because identifying meaningful failure cases often requires substantial human effort to design realistic

safetyarxiv-cs-cl
19 May 2026
Safety

Real2Sim via Active Perception with Behavior Trees Automatically Generated by VLMs

DGX agent

arXiv:2601.08454v2 Announce Type: replace Abstract: Constructing physically accurate simulation environments (Real2Sim) traditionally relies on manual system identification or rigid, exhaustive explor

safetyarxiv-cs-ro
19 May 2026
Safety

SuReNav: Superpixel Graph-based Constraint Relaxation for Navigation in Over-constrained Environments

DGX agent

arXiv:2602.06807v2 Announce Type: replace-cross Abstract: We address the over-constrained planning problem in semi-static environments. The planning objective is to find a best-effort solution that av

safetyarxiv-cs-ai
19 May 2026
Safety

TClone: Low-Latency Forking of Live GUI Environments for Computer-Use Agents

DGX agent

arXiv:2605.17320v1 Announce Type: cross Abstract: Computer-use agents increasingly operate inside live personal workspaces, where their actions can modify files, applications, GUI state, credentials,

safetyarxiv-cs-ai
19 May 2026
Safety

Temporal Task Diversity: Inductive Biases Under Non-Stationarity in Synthetic Sequence Modelling

DGX agent

arXiv:2605.18281v1 Announce Type: new Abstract: Modern deep learning science often assumes that neural networks learn from a fixed data distribution. However, many practically important learning probl

safetyarxiv-cs-lg
19 May 2026
Safety

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

DGX agent

arXiv:2605.17480v1 Announce Type: new Abstract: Multi-agent systems extend large language models (LLMs) by decomposing tasks among specialized agents, but their distributed decision process creates ne

safetyarxiv-cs-ai
19 May 2026
Safety

Transformer-Based MCS Prediction for 5G Multicast-Broadcast Services (MBS)

DGX agent

arXiv:2605.16735v1 Announce Type: cross Abstract: The deployment of 5G Multicast-Broadcast Services (MBS) is emerging as a critical technology for spectral-efficient UHD content delivery and serving a

safetyarxiv-cs-lg
19 May 2026
Safety

Unleashing the Potential of Diffusion Models for End-to-End Autonomous Driving

DGX agent

arXiv:2602.22801v2 Announce Type: replace-cross Abstract: Diffusion models have become a popular choice for decision-making tasks in robotics, and more recently, are also being considered for solving

safetyarxiv-cs-ai
19 May 2026
Safety

What we find most useful about CNA is that the intervention is simple yet powerful. The steering is a multiplicative ablation on a sparse se…

DGX agent

What we find most useful about CNA is that the intervention is simple yet powerful. The steering is a multiplicative ablation on a sparse set of MLP neurons, which makes CNA a clean addition on top of

safetynous-research--x
19 May 2026
Safety

AI Consciousness and Existential Risk

DGX agent

arXiv:2511.19115v2 Announce Type: replace Abstract: In AI, the existential risk denotes the hypothetical threat posed by an artificial system that would possess both the capability and the objective,

safetyarxiv-cs-ai
18 May 2026
Safety

ASRU: Activation Steering Meets Reinforcement Unlearning for Multimodal Large Language Models

DGX agent

arXiv:2605.15687v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) may memorize sensitive cross-modal information during pretraining, making machine unlearning (MU) crucial. Ex

safetyarxiv-cs-ai
18 May 2026
Safety

Constrained MPC-Based Motion Planning for Morphing Quadrotors in Ultra-Narrow Passages under Limited Perception

DGX agent

arXiv:2605.15999v1 Announce Type: new Abstract: This paper introduces a motion planning framework to plan morphology and trajectory for morphing quadrotors under extremely constrained environments. We

safetyarxiv-cs-ro
18 May 2026
Safety

CTF4Nuclear: Common Task Framework for Nuclear Fission and Fusion Models

DGX agent

arXiv:2605.15549v1 Announce Type: cross Abstract: The demand for clean energy is ever increasing, with new nuclear technologies presenting a complementary solution to renewable energies. However, desi

safetyarxiv-cs-ai
18 May 2026
Safety

Driving Through the Network: Performance and Workload Under Latency and Video Impairment

DGX agent

arXiv:2605.15952v1 Announce Type: cross Abstract: Teleoperation promises to extend the operational envelope of automated vehicles, yet it critically depends on network latency and video quality. We re

safetyarxiv-cs-ro
18 May 2026
Safety

Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems

DGX agent

arXiv:2605.16198v1 Announce Type: new Abstract: We examine one particular dimension of AI governance: how to monitor and audit AI-enabled products and services throughout the AI development lifecycle,

safetyarxiv-cs-ai
18 May 2026
Safety

How Google Does It: Fleet-wide, large-scale A/B experimentation

DGX agent

When most people think of A/B experimentation, they think of button colors, landing page layouts, or checkout flows. At Google, many fundamental infrastructure improvements also need the rigor of A/B

safetygoogle-cloud-ai
18 May 2026
Safety

Learning in Structured Stackelberg Games

DGX agent

arXiv:2504.09006v4 Announce Type: replace-cross Abstract: We initiate the study of structured Stackelberg games, a novel form of strategic interaction between a leader and a follower where contextual

safetyarxiv-cs-lg
18 May 2026
Safety

OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation

DGX agent

arXiv:2605.15971v1 Announce Type: new Abstract: While reinforcement learning (RL) enables robots to acquire skills autonomously, its real-world deployment is severely limited by inefficient and unsafe

safetyarxiv-cs-ro
18 May 2026
Safety

Quantum Artificial Intelligence for Mission-Critical Systems: Foundations, Architectural Elements, and Future Directions

DGX agent

arXiv:2511.09884v2 Announce Type: replace Abstract: Mission critical (MC) applications such as defense operations, energy management, cybersecurity, and aerospace control require reliable, determinist

safetyarxiv-cs-ai
18 May 2026
Safety

SAFE Quantum Machine Learning with Variational Quantum Classifiers

DGX agent

arXiv:2605.16067v1 Announce Type: new Abstract: We propose a variational quantum classifier operating on high dimensional deep representations via amplitude encoding, stabilized by a learnable classic

safetyarxiv-cs-lg
18 May 2026
Safety

Towards Trustworthy and Explainable AI for Perception Models: From Concept to Prototype Vehicle Deployment

DGX agent

arXiv:2605.16087v1 Announce Type: cross Abstract: Deep Neural Networks have become the dominant solution for Autonomous Driving perception, but their opacity conflicts with emerging Trustworthy AI gui

safetyarxiv-cs-ai
18 May 2026
Safety

What I am about to describe ain’t AGI; it’s a sign of a trillion dollar trainwreck. If I had told you in 2022 that the 2026 version of GPT (…

DGX agent

What I am about to describe ain’t AGI; it’s a sign of a trillion dollar trainwreck. If I had told you in 2022 that the 2026 version of GPT (which by the way would only be GPT 5.5 and not GPT-6 or 7 li

safetygary-marcus--x
17 May 2026
Safety

A Regret Perspective on Online Multiple Testing

DGX agent

arXiv:2605.13916v1 Announce Type: cross Abstract: Online Multiple Testing (OMT), a fundamental pillar of sequential statistical inference, traditionally evaluates the False Discovery Rate (FDR) and st

safetyarxiv-cs-ai
15 May 2026
Safety

Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model

DGX agent

arXiv:2605.14723v1 Announce Type: new Abstract: Sepsis management in the ICU requires sequential treatment decisions under rapidly evolving patient physiology. Although large language models (LLMs) en

safetyarxiv-cs-ai
15 May 2026
Safety

Behavioral Data-Driven Optimal Trajectory Generation for Rotary Cranes

DGX agent

arXiv:2605.14944v1 Announce Type: new Abstract: With the growth of the construction industry and the global shortage of skilled labor, the automation of crane control has become increasingly important

safetyarxiv-cs-ro
15 May 2026
Safety

DIVER: Reinforced Diffusion Breaks Imitation Bottlenecks in End-to-End Autonomous Driving

DGX agent

arXiv:2507.04049v4 Announce Type: replace Abstract: Most end-to-end autonomous driving methods rely on imitation learning from single expert demonstrations, often leading to conservative and homogeneo

safetyarxiv-cs-cv
15 May 2026
Safety

Fair and Calibrated Toxicity Detection with Robust Training and Abstention

DGX agent

arXiv:2605.14074v1 Announce Type: new Abstract: Fairness in toxicity classification involves three integrated axes: ranking, calibration, and abstention. Training-time interventions and post-hoc safet

safetyarxiv-cs-lg
15 May 2026
Safety

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving

DGX agent

arXiv:2605.14201v1 Announce Type: cross Abstract: Vision-language-action (VLA) models are effective as end-to-end motion planners, but can be brittle when evaluated in closed-loop settings due to bein

safetyarxiv-cs-cv
15 May 2026
Safety

MetaBackdoor: Exploiting Positional Encoding as a Backdoor Attack Surface in LLMs

DGX agent

arXiv:2605.15172v1 Announce Type: cross Abstract: Backdoor attacks pose a serious security threat to large language models (LLMs), which are increasingly deployed as general-purpose assistants in safe

safetyarxiv-cs-cl
15 May 2026
Safety

Novel Dynamic Batch-Sensitive Adam Optimiser for Vehicular Accident Injury Severity Prediction

DGX agent

arXiv:2605.15083v1 Announce Type: cross Abstract: The choice of optimiser is important in deep learning, as it strongly influences model efficiency and speed of convergence. However, many commonly use

safetyarxiv-cs-ai
15 May 2026
Safety

Quantifying and Mitigating Premature Closure in Frontier LLMs

DGX agent

arXiv:2605.15000v1 Announce Type: cross Abstract: Premature closure, or committing to a conclusion before sufficient information is available, is a recognized contributor to diagnostic error but remai

safetyarxiv-cs-ai
15 May 2026
Safety

Viverra: Text-to-Code with Guarantees

DGX agent

arXiv:2605.14972v1 Announce Type: cross Abstract: A fundamental limitation of Text-to-Code is that no guarantee can be obtained about the correctness of the generated code. Therefore, to ensure its co

safetyarxiv-cs-ai
15 May 2026
Safety

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

DGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

safetyarxiv-cs-lg
14 May 2026
Safety

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

DGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

safetyarxiv-cs-ai
14 May 2026
Safety

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

DGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

safetyarxiv-cs-ai
14 May 2026
Safety

DiffusionHijack: Supply-Chain PRNG Backdoor Attack on Diffusion Models and Quantum Random Number Defense

DGX agent

arXiv:2605.13115v1 Announce Type: cross Abstract: Diffusion models depend on pseudo-random number generators (PRNGs) for latent noise sampling. We present DiffusionHijack, a supply-chain backdoor atta

safetyarxiv-cs-lg
14 May 2026
Safety

Interesting position paper on agentic AI as a foreseeable pathway to AGI. (bookmark it) There has been strong debate on whether a larger sin…

DGX agent

Interesting position paper on agentic AI as a foreseeable pathway to AGI. (bookmark it) There has been strong debate on whether a larger single model get us there or a multi-agent system. The authors

safetydair-ai--x
14 May 2026
Safety

Loiter UAV Reinsertion Guidance for Fixed-wing UAV Corridors

DGX agent

arXiv:2605.13822v1 Announce Type: new Abstract: This paper considers fixed-wing unmanned aerial vehicle (UAV) corridors comprising a main lane, a circular loiter lane for managing traffic congestion,

safetyarxiv-cs-ro
14 May 2026
Model Releases

Neurosymbolic Auditing of Natural-Language Software Requirements

DGX agent

arXiv:2605.13817v1 Announce Type: cross Abstract: Natural-language software requirements are often ambiguous, inconsistent, and underspecified; in safety-critical domains, these defects propagate into

model-releasesarxiv-cs-ai
14 May 2026
Safety

No Attack Required: Semantic Fuzzing for Specification Violations in Agent Skills

DGX agent

arXiv:2605.13044v1 Announce Type: cross Abstract: LLM-powered agents can silently delete documents, leak credentials, or transfer funds on a routine user request, not because the agent was attacked, b

safetyarxiv-cs-ai
14 May 2026
Safety

Runtime Monitoring of Perception-Based Autonomous Systems via Embedding Temporal Logic

DGX agent

arXiv:2605.12651v1 Announce Type: new Abstract: Runtime monitoring of autonomous systems traditionally relies on mapping continuous sensor observations to discrete logical propositions defined over lo

safetyarxiv-cs-lg
14 May 2026
← Previous
1…5556575859…300
Next →