AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,356 results
Safety

Governance by Design: Architecting Agentic AI for Organizational Learning and Scalable Autonomy

DGX agent

arXiv:2605.20210v1 Announce Type: cross Abstract: Agentic AI systems - systems that can pursue goals through multi-step planning and tool-mediated action with limited direct supervision - are moving f

safetyarxiv-cs-ai
22 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

How can reasoning capability empower the AI copilot robot in endoscopic surgery

DGX agent

arXiv:2605.22322v1 Announce Type: new Abstract: Reasoning capability has significantly advanced complex logical inference and robotic decision-making in general domains. However, its potential in the

safetyarxiv-cs-ro
22 May 2026
Safety

Making the Discrete Continuous: Synthetic RAW Augmentations for Fine-Grained Evaluation of Person Detection Performance in Low Light

DGX agent

arXiv:2605.22455v1 Announce Type: new Abstract: Real-world deployment of AI vision models is both fueled and limited by the data available for training and testing. Real datasets are sparse and uneven

safetyarxiv-cs-cv
22 May 2026
Safety

Open-source LLMs administer maximum electric shocks in a Milgram-like obedience experiment

DGX agent

arXiv:2605.21401v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that make sequences of decisions over extended interactions in high-stakes

safetyarxiv-cs-ai
22 May 2026
Safety

Parallel OctoMapping: A Scalable Framework for Enhanced Path Planning in Autonomous Navigation

DGX agent

arXiv:2603.22508v2 Announce Type: replace Abstract: Mapping is essential in robotics and autonomous systems because it provides the spatial foundation for path planning. Efficient mapping enables plan

safetyarxiv-cs-ro
22 May 2026
Safety

Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO (Washington Post)

DGX agent

Washington Post: Sources: last-minute calls with Elon Musk, Mark Zuckerberg, David Sacks, and others helped persuade President Trump not to sign the highly anticipated AI EO — Industry leaders warned

safetytechmeme
22 May 2026
Safety

update @thestalwart found more recent models less vulnerable. would be good to do a broad study of this.

DGX agent

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models woul

safetygary-marcus--x
22 May 2026
Safety

AI-Assisted Competency Assessment from Egocentric Video in Simulation-Based Nursing Education

DGX agent

arXiv:2605.20233v1 Announce Type: new Abstract: Assessing learner competency in clinical simulation requires expert observation that is time-intensive, difficult to scale, and subject to inter-rater v

safetyarxiv-cs-cv
21 May 2026
Safety

AIMBio-Mat: An AI-Native FAIR Platform for Closed-Loop Materials Discovery and Biomedical Translation

DGX agent

arXiv:2605.21083v1 Announce Type: cross Abstract: Materials discovery and biomedical translation increasingly require models that can reason across composition, processing, structure, biological respo

safetyarxiv-cs-lg
21 May 2026
Safety

CHEM: Estimating and Understanding Hallucinations in Deep Learning for Image Processing

DGX agent

arXiv:2512.09806v2 Announce Type: replace Abstract: Deep learning-based methods have recently achieved significant success in image reconstruction problems. However, challenges have emerged, as these

safetyarxiv-cs-cv
21 May 2026
Safety

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models

DGX agent

arXiv:2605.20591v1 Announce Type: new Abstract: Medical large language models (LLMs), including custom medical GPTs (MedGPTs) and open-source models, are increasingly deployed on web platforms to prov

safetyarxiv-cs-cl
21 May 2026
Safety

FineVision: Open Data Is All You Need

DGX agent

arXiv:2510.17269v2 Announce Type: replace Abstract: The advancement of vision-language models (VLMs) is hampered by a fragmented landscape of inconsistent and contaminated public datasets. We introduc

safetyarxiv-cs-cv
21 May 2026
Safety

GAMR: Geometric-Aware Manifold Regularization with Virtual Outlier Synthesis for Learning with Noisy Labels

DGX agent

arXiv:2605.20727v1 Announce Type: new Abstract: Deep neural networks (DNNs) experience significant performance degradation when processing noisy labels, primarily due to overfitting on mislabeled data

safetyarxiv-cs-cv
21 May 2026
Safety

GraphDiffMed: Knowledge-Constrained Differential Attention with Pharmacological Graph Priors for Medication Recommendation

DGX agent

arXiv:2605.20188v1 Announce Type: new Abstract: Recommending safe and effective medication combinations from electronic health records (EHRs) is a core clinical AI problem, yet it remains difficult be

safetyarxiv-cs-lg
21 May 2026
Safety

Lost in Fog: Sensor Perturbations Expose Reasoning Fragility in Driving VLAs

DGX agent

arXiv:2605.21446v1 Announce Type: new Abstract: Interpretable autonomous driving planners depend not only on generating explanations, but also on those explanations remaining reliable under real-world

safetyarxiv-cs-ro
21 May 2026
Safety

Mechanistic Interpretability for Learning Assurance of a Vision-Based Landing System

DGX agent

arXiv:2605.20607v1 Announce Type: cross Abstract: EASA's learning-assurance guidance requires data-driven aviation systems to build and monitor their own situation representation, yet for neural netwo

safetyarxiv-cs-cv
21 May 2026
Safety

Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty

DGX agent

arXiv:2605.20255v1 Announce Type: new Abstract: Simulation-based testing of self-driving cars (SDCs) typically relies on scripted or simplified pedestrian models that do not capture the heterogeneity

safetyarxiv-cs-lg
21 May 2026
Safety

Preference-aware Influence-function-based Data Selection Method for Efficient Fine-Tuning

DGX agent

arXiv:2605.21422v1 Announce Type: new Abstract: As LLMs continue to scale, improving training efficiency increasingly depends on using data more effectively. Data selection addresses this problem by a

safetyarxiv-cs-lg
21 May 2026
Safety

Quadratic Characterizations for Reachability Analysis of Neural Networks

DGX agent

arXiv:2605.20482v1 Announce Type: new Abstract: Quadratic constraints (QCs) are widely used to characterize nonlinearities and uncertainties, but generic analytical characterizations can be conservati

safetyarxiv-cs-lg
21 May 2026
Safety

Spacetime Optimal-Transport Attention for Visuo-Haptic Imitation Learning of Contact-Rich Manipulation

DGX agent

arXiv:2605.20433v1 Announce Type: new Abstract: Contact-rich manipulation tasks such as tight-clearance insertion, connector mating, polishing, and surface-conforming wiping remain difficult for data-

safetyarxiv-cs-ro
21 May 2026
Safety

SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary

DGX agent

arXiv:2605.21132v1 Announce Type: new Abstract: Understanding surgical workflow in real time is fundamental for intelligent surgical embodiment, where AI systems continuously perceive and respond as s

safetyarxiv-cs-cv
21 May 2026
Safety

Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs

DGX agent

arXiv:2605.20641v1 Announce Type: cross Abstract: Inference optimization is a vital technique for deploying LLMs at scale. Compilation is the most widely adopted optimization technique for LLMs. While

safetyarxiv-cs-lg
21 May 2026
Safety

When AI Gets it Wrong: Reliability and Risk in AI-Assisted Medication Decision Systems

DGX agent

arXiv:2604.01449v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) systems are increasingly integrated into healthcare and pharmacy workflows, supporting tasks such as medication r

safetyarxiv-cs-lg
21 May 2026
Safety

ARC-RL: A Reinforcement Learning Playground Inspired by ARC Raiders

DGX agent

arXiv:2605.19503v1 Announce Type: cross Abstract: Reinforcement learning for legged locomotion has matured into a stack of multi-component reward functions and physics-engine benchmarks whose morpholo

safetyarxiv-cs-ai
20 May 2026
Safety

Attention-Guided Reward for Reinforcement Learning-based Jailbreak against Large Reasoning Models

DGX agent

arXiv:2605.19485v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) have demonstrated remarkable capabilities in solving complex problems by generating structured, step-by-step reasoning con

safetyarxiv-cs-ai
20 May 2026
Safety

Backtracking When It Strays: Mitigating Dual Exposure Biases in LLM Reasoning Distillation

DGX agent

arXiv:2605.19433v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable success in complex reasoning tasks via long chain-of-thought (CoT), yet their immense computatio

safetyarxiv-cs-ai
20 May 2026
Safety

Be Kind, Rewrite: Benign Projections via Rewriting Defend Against LLM Data Poisoning Attacks

DGX agent

arXiv:2605.19147v1 Announce Type: cross Abstract: Large language models (LLMs) are highly susceptible to backdoor attacks (BAs), wherein training samples are poisoned using trigger-based harmful conte

safetyarxiv-cs-ai
20 May 2026
Model Releases

Beyond Imitation: Learning Safe End-to-End Autonomous Driving from Hard Negatives

DGX agent

arXiv:2605.19771v1 Announce Type: cross Abstract: Existing imitation learning methods for end-to-end autonomous driving predominantly learn from successful demonstrations by minimizing geometric devia

model-releasesarxiv-cs-cv
20 May 2026
Safety

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced wit…

DGX agent

⚠️👇 🚨Breaking ⚠️ If we can’t make AI agents follow rules, we are screwed. New study from METR reports that “when the agents were faced with hard tasks, they routinely violated constraints” This—routin

safetygary-marcus--x
20 May 2026
Safety

CADENet: Condition-Adaptive Asynchronous Dual-Stream Enhancement Network for Adverse Weather Perception in Autonomous Driving

DGX agent

arXiv:2605.19837v1 Announce Type: cross Abstract: Adverse weather (rain, fog, sand, and snow) degrades camera-based object detection in autonomous vehicles. Existing enhancement-then-detect approaches

safetyarxiv-cs-ai
20 May 2026
Safety

CEPO: RLVR Self-Distillation using Contrastive Evidence Policy Optimization

DGX agent

arXiv:2605.19436v1 Announce Type: cross Abstract: When a model produces a correct solution under reinforcement learning with verifiable rewards (RLVR), every token receives the same reward signal rega

safetyarxiv-cs-cl
20 May 2026
Safety

Compliant Explicit Reference Governor for Contact Friendly Robotic Manipulators

DGX agent

arXiv:2504.09188v2 Announce Type: replace Abstract: This paper introduces the Compliant Explicit Reference Governor (CERG), a modular reference management system that enables robots to interact physic

safetyarxiv-cs-ro
20 May 2026
Model Releases

Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations

DGX agent

arXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga

model-releasesarxiv-cs-ai
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Model Releases

Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support

DGX agent

arXiv:2605.19826v1 Announce Type: new Abstract: Operators of safety-critical industrial processes increasingly rely on digital twins to screen control interventions, but such simulators rarely carry c

model-releasesarxiv-cs-ai
20 May 2026
Safety

Graph Neural Planning and Predictive Control for Multi-Robot Communication-Constrained Unlabeled Motion Planning

DGX agent

arXiv:2605.19209v1 Announce Type: new Abstract: The multi-robot unlabeled motion planning problem of concurrently assigning robots to goals and generating safe trajectories is central in many collabor

safetyarxiv-cs-ro
20 May 2026
Safety

Hamilton--Jacobi Reachability for Spacecraft Collision Avoidance

DGX agent

arXiv:2605.20138v1 Announce Type: new Abstract: This article presents a Hamilton--Jacobi (HJ) reachability framework for a two--satellite collision avoidance problem operating in the same circular orb

safetyarxiv-cs-ro
20 May 2026
Local Ai

KIO-planner: Attention-Guided Single-Stage Motion Planning with Dual Mapping for UAV Navigation

DGX agent

arXiv:2605.19703v1 Announce Type: new Abstract: Autonomous UAV flight in confined, wall-dense environments requires low-latency and reliable motion planning under strict safety constraints. Traditiona

local-aiarxiv-cs-ro
20 May 2026
Safety

Knowing When Not to Predict: Self Supervised Learning and Abstention for Safer DR Screening

DGX agent

arXiv:2605.19133v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is now a standard way to pretrain medical image models, but performance is still mostly judged by downstream accuracy.

safetyarxiv-cs-ai
20 May 2026
Safety

Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering

DGX agent

arXiv:2605.19220v1 Announce Type: cross Abstract: Uncertainty Quantification (UQ) is widely regarded as the primary safeguard for deploying Large Language Models (LLMs) in high-stakes domains. However

safetyarxiv-cs-ai
20 May 2026
Safety

Probabilistic Recursively Feasible Motion Planning Under Uncertain Environments

DGX agent

arXiv:2605.19015v1 Announce Type: cross Abstract: Safe motion planning in uncertain, time-varying environments is challenging because the safe region can change unpredictably across planning steps, of

safetyarxiv-cs-ro
20 May 2026
Safety

Pseudocode-Guided Structured Reasoning for Automating Reliable Inference in Vision-Language Models

DGX agent

arXiv:2605.19663v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are becoming the cornerstone of high-level reasoning for robotic automation, enabling robots to parse natural language com

safetyarxiv-cs-ai
20 May 2026
Safety

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

DGX agent

arXiv:2605.19940v1 Announce Type: new Abstract: Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cum

safetyarxiv-cs-ai
20 May 2026
Safety

Augmenting Human Evaluation with LLM Judges: How Many Human Reviews Do You Need?

DGX agent

arXiv:2605.16354v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as automated evaluators of AI systems, including in high-stakes applications. In this role, LLMs ar

safetyarxiv-cs-ai
19 May 2026
Safety

Code as Agent Harness

DGX agent

arXiv:2605.18747v1 Announce Type: cross Abstract: Recent large language models (LLMs) have demonstrated strong capabilities in understanding and generating code, from competitive programming to reposi

safetyarxiv-cs-ai
19 May 2026
Safety

Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems

DGX agent

arXiv:2605.17909v1 Announce Type: new Abstract: As autonomous agentic systems scale across regulated critical infrastructures, the lack of mechanistic, hardware-rooted enforcement for high-frequency p

safetyarxiv-cs-ai
19 May 2026
Safety

Event-Grounded Sparse Autoencoders for Vision-Language-Action Policies

DGX agent

arXiv:2605.17204v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies translate language and visual inputs into robot actions, where their hidden representations directly shape close

safetyarxiv-cs-ai
19 May 2026
Safety

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

DGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

safetyarxiv-cs-ai
19 May 2026
← Previous
1…5455565758…300
Next →