AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

OGM-CBF: Occupancy Grid Map-based Control Barrier Function for Safe Mobile Robot Control with Memory of out of View Obstacles

DGX agent

arXiv:2405.10703v4 Announce Type: replace Abstract: Safe control in unknown environments is a key challenge in mobile robotics. Control Barrier Functions (CBFs) provide a principled framework for guar

safetyarxiv-cs-ro
30 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

SPACE: Swarm Pheromone Fields for Adaptive Collision-Aware Exploration

DGX agent

arXiv:2606.29372v1 Announce Type: new Abstract: Massive robot swarms can explore unknown environments quickly, but adding robots eventually stops helping. Doorways and dense traffic create congestion,

safetyarxiv-cs-ro
30 Jun 2026
Safety

Low-Agreeableness Persona Conditioning for Safe LLM Fine-Tuning

DGX agent

arXiv:2606.27709v1 Announce Type: cross Abstract: Recent work has shown that fine-tuning large language models (LLMs) for social warmth degrades factual reliability and increases sycophancy. We invest

safetyarxiv-cs-ai
29 Jun 2026
Safety

Human-AI Complementarity: A Goal for Amplified Oversight

DGX agent

arXiv:2510.26518v2 Announce Type: replace Abstract: Human feedback is critical for aligning AI systems to human values. As AI capabilities improve and AI is used to tackle more challenging tasks, veri

safetyarxiv-cs-ai
26 Jun 2026
Safety

Proposal-Conditioned Latent Diffusion for Closed-Loop Traffic Scenario Generation

DGX agent

arXiv:2606.27123v1 Announce Type: cross Abstract: Closed-loop traffic simulation remains challenging because it must generate interactive multi-agent behaviors that are scene-consistent and controllab

safetyarxiv-cs-cv
26 Jun 2026
Safety

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks

DGX agent

arXiv:2606.27147v1 Announce Type: cross Abstract: Unlike diffusion-based models that operate in continuous latent spaces, autoregressive unified multimodal models produce images by sequentially predic

safetyarxiv-cs-ai
26 Jun 2026
Safety

G2DP: Diffusion Planning with Spatio-Temporal Grid Guidance

DGX agent

arXiv:2606.26017v1 Announce Type: new Abstract: In autonomous driving, diffusion-based planners have emerged as a promising paradigm for robust motion planning in dense and interactive traffic, as the

safetyarxiv-cs-ro
25 Jun 2026
Safety

MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction

DGX agent

arXiv:2606.25651v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare settings, accurate error detection and correction in generated or existing text

safetyarxiv-cs-cl
25 Jun 2026
Safety

A Geometry-Informed Computer Vision Method for Detecting and Examining Overtaking Vehicles From A Bicycle

DGX agent

arXiv:2606.23699v1 Announce Type: new Abstract: Instrumented bicycle studies have produced direct field evidence on vehicle passing behavior, but extracting overtaking events from continuous rear-faci

safetyarxiv-cs-cv
24 Jun 2026
Safety

FlowR2A: Learning Reward-to-Action Distribution for Multimodal Driving Planning

DGX agent

arXiv:2606.24231v1 Announce Type: new Abstract: Multimodal driving planning faces a long-standing tension between two paradigms: scoring-based methods benefit from dense reward supervision but are con

safetyarxiv-cs-ai
24 Jun 2026
Safety

HiPath: Hierarchical Vision-Language Alignment for Structured Pathology Report Prediction

DGX agent

arXiv:2603.19957v2 Announce Type: replace-cross Abstract: Pathology reports are structured, multi-granular documents encoding diagnostic conclusions, histological grades, and ancillary test results ac

safetyarxiv-cs-ai
24 Jun 2026
Safety

PHANTOM: A Large-Scale Dataset of Multimodal Adversarial Attacks for Vision-Language Models

DGX agent

arXiv:2606.24388v1 Announce Type: new Abstract: We introduce a large-scale, open-source dataset of pre-generated adversarial attacks for vision-language models (VLMs). The dataset is designed to be di

safetyarxiv-cs-ai
24 Jun 2026
Safety

BiliVLA: Scene-Aware Vision-Language-Action Model with Reinforcement Learning for Autonomous Biliary Endoscopic Navigation

DGX agent

arXiv:2606.23531v1 Announce Type: new Abstract: Endoscopic retrograde cholangiopancreatography (ERCP) demands precise endoscopic navigation and stable biliary cannulation within a narrow monocular fie

safetyarxiv-cs-ro
23 Jun 2026
Safety

Do Activation Monitors Survive Model Updates? Benchmarking, Predicting, and Repairing Activation-Monitor Staleness

DGX agent

arXiv:2606.15980v2 Announce Type: replace Abstract: Activation monitors -- lightweight probes trained on a language model's internal representations -- are an increasingly common layer in deployment s

safetyarxiv-cs-lg
23 Jun 2026
Safety

Expert Consensus on Criteria for the Automated Assessment of Laparoscopic Camera Navigation

DGX agent

arXiv:2606.23131v1 Announce Type: new Abstract: Background: Laparoscopic camera navigation (LCN) is a critical skill, yet its current assessment typically relies on manual rating systems which are tim

safetyarxiv-cs-cv
23 Jun 2026
Safety

Intent-Handover: Grounding Language in Human-Usage Regions for Trustworthy Robot-to-Human Handovers

DGX agent

arXiv:2503.03579v2 Announce Type: replace-cross Abstract: Spoken instructions in robot-to-human handovers may specify either an object ('the cup') or an intended use ('pour water'); in both cases, suc

safetyarxiv-cs-lg
23 Jun 2026
Safety

One Image is All You Need: Agentic One-Shot Image Generation via Text-Based World Models for Long-Tail Spatial Perception

DGX agent

arXiv:2606.20764v1 Announce Type: new Abstract: Reliable spatial decision automation, such as autonomous driving and maritime surveillance, critically depends on robust visual perception. However, rea

safetyarxiv-cs-cv
23 Jun 2026
Safety

SkillHarness: Harnessing Safe Skills for Computer-Use Agents

DGX agent

arXiv:2606.20636v1 Announce Type: cross Abstract: Computer-Use Agents (CUAs) are increasingly deployed in dynamic interactive environments, creating a growing need for continual skill learning during

safetyarxiv-cs-lg
23 Jun 2026
Safety

AerialClaw: An Open-Source Framework for LLM-Driven Autonomous Aerial Agents

DGX agent

arXiv:2606.12142v1 Announce Type: cross Abstract: Unmanned aerial vehicles (UAVs) are increasingly used in inspection, search and rescue, environmental monitoring, and emergency response. However, mos

safetyarxiv-cs-cv
11 Jun 2026
Safety

Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code

DGX agent

arXiv:2606.11817v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, raising concerns that they may be misused to produce malicious code. Meanwhile

safetyarxiv-cs-ai
11 Jun 2026
Safety

UGV-Conditioned Multi-UAV Informative Planning on a Shared Exposure Belief

DGX agent

arXiv:2606.12306v1 Announce Type: new Abstract: Safe ground navigation in large, threat-augmented environments requires aerial support that actively reduces the risks that a ground vehicle faces along

safetyarxiv-cs-ro
11 Jun 2026
Safety

EM-Fall: Embodied mmWave Sensing for Day-and-Night Fall Detection on Humanoid Robots

DGX agent

arXiv:2606.11109v1 Announce Type: new Abstract: Falls are one of the leading causes of injury and hospitalization among elderly individuals, making reliable fall awareness an essential capability for

safetyarxiv-cs-ro
10 Jun 2026
Safety

AI Assurance in UK Defence: Challenges in Operationalising JSP 936

DGX agent

arXiv:2606.09414v1 Announce Type: cross Abstract: This report examines practical challenges in operationalising JSP 936 Part 1 for AI assurance in UK Defence. Using a structured interpretive review of

safetyarxiv-cs-ai
9 Jun 2026
Safety

Autonomous Incident Resolution at Hyperscale: An Agentic AI Architecture for Network Operations

DGX agent

arXiv:2606.09122v1 Announce Type: cross Abstract: Cloud network infrastructure at hyperscale presents unique operational challenges where traditional human-driven incident response cannot keep pace wi

safetyarxiv-cs-ai
9 Jun 2026
Safety

Can the Environment Speak for Itself? T^{2}-GRPO: A Turn-Trajectory Group Relative Policy Optimization for Caregiver Agents

DGX agent

arXiv:2606.08875v1 Announce Type: new Abstract: Optimizing large language models (LLMs) for long-horizon caregiver agents requires balancing delayed task objectives with immediate environment dynamics

safetyarxiv-cs-ai
9 Jun 2026
Safety

Fast LLM-Based Semantic Filtering: From a Unified Framework to an Adaptive Two-Phase Method

DGX agent

arXiv:2606.08090v1 Announce Type: cross Abstract: Evaluating a natural-language yes/no predicate over a document corpus under an accuracy target - the semantic filter - is a cornerstone of LLM-based d

safetyarxiv-cs-ai
9 Jun 2026
Safety

MC-CPO: Mastery-Conditioned Constrained Policy Optimization for Pedagogically Safe Intelligent Tutoring Systems

DGX agent

arXiv:2604.04251v2 Announce Type: replace Abstract: Intelligent tutoring systems increasingly rely on reinforcement learning to personalise instruction, yet optimising for observable engagement signal

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Safe-RULE: Safe Reinforcement UnLEarning

DGX agent

arXiv:2606.09559v1 Announce Type: cross Abstract: Offline safe reinforcement learning (Safe RL) enables policy learning without online interactions, making it suitable for safety-critical systems such

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Semantic Quorum Assurance: Collective Certification for Non-Deterministic AI Infrastructure

DGX agent

arXiv:2606.08021v1 Announce Type: cross Abstract: As large language model (LLM) agents are integrated into autonomous cloud operations, distributed systems face a semantic reliability problem: propose

safetyarxiv-cs-ai
9 Jun 2026
Safety

Workflow-to-Skill: Skill Creation via Routing-Workflow-Semantics-Attachments Decomposition

DGX agent

arXiv:2606.06893v1 Announce Type: new Abstract: Large language model agents increasingly rely on Skills to encode procedural knowledge, yet high-quality Skills remain costly to hand-write. This paper

safetyarxiv-cs-ai
8 Jun 2026
Safety

UNIVID: Unified Vision-Language Model for Video Moderation

DGX agent

arXiv:2606.05748v1 Announce Type: cross Abstract: Global-scale video moderation faces a dual challenge: the need for fine-grained multi-modal reasoning and the demand for interpretable outputs to supp

safetyarxiv-cs-cl
5 Jun 2026
Safety

Distribution-Free Risk-Aware Planning and Control Under Uncertainty Using Conformal Spectral Risk Control

DGX agent

arXiv:2606.04185v1 Announce Type: new Abstract: Safe navigation in dynamic and uncertain environments often relies on accurate estimation of, or assumptions about, the true underlying uncertainty. How

safetyarxiv-cs-ro
4 Jun 2026
Safety

From Agent Traces to Trust: Evidence Tracing and Execution Provenance in LLM Agents

DGX agent

arXiv:2606.04990v1 Announce Type: cross Abstract: Large language model (LLM)-based agents increasingly solve complex tasks by interacting with external tools, retrieval systems, memory modules, enviro

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

DGX agent

arXiv:2603.03205v2 Announce Type: replace Abstract: Agentic language models operate in a fundamentally different safety regime than chat models: they must plan, call tools, and execute long-horizon ac

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Reasoning over Boundaries: Enhancing Specification Alignment via Test-time Deliberation

DGX agent

arXiv:2509.14760v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in diverse real-world scenarios, each governed by bespoke behavioral and safety specifications

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

RSC: Decentralized Rigid Formation Flocking for Large-Scale Swarms via Hybrid Predictive Control and Online Reconfiguration

DGX agent

arXiv:2606.04248v1 Announce Type: new Abstract: Decentralized rigid formation flocking requires a swarm of autonomous agents to maintain a predetermined geometric configuration while moving, relying s

safetyarxiv-cs-ro
4 Jun 2026
Safety

Bridging Predictive Uncertainty and Safe Action: Sample-Conditioned Differentiable Planning for Autonomous Driving

DGX agent

arXiv:2606.03296v1 Announce Type: new Abstract: Complex, dynamic, and interactive driving environments pose significant challenges for autonomous driving, primarily due to the pervasive uncertainty of

safetyarxiv-cs-ro
3 Jun 2026
Local Ai

DDOR: Delta Debugging for Explainable Overrefusal Testing and Repair

DGX agent

arXiv:2606.03601v1 Announce Type: cross Abstract: While safety alignment and guardrails help large language models (LLMs) avoid harmful outputs, they can also induce overrefusal, i.e., unwarranted rej

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Latent Activation Editing: Inference-Time Refinement of Learned Policies for Safer Multirobot Navigation

DGX agent

arXiv:2509.20623v2 Announce Type: replace Abstract: Reinforcement learning has enabled significant progress in complex domains such as coordinating and navigating multiple quadrotors. However, even we

safetyarxiv-cs-ro
3 Jun 2026
Safety

Towards a Science of AI Agent Reliability

DGX agent

arXiv:2602.16666v3 Announce Type: replace Abstract: AI agents are increasingly deployed to execute important tasks. While rising accuracy scores on standard benchmarks suggest rapid progress, many age

safetyarxiv-cs-ai
3 Jun 2026
Safety

When Models Refuse: Political Steerability and Feature Richness as Measures of Ideological Depth

DGX agent

arXiv:2508.21448v3 Announce Type: replace Abstract: Large language models (LLMs) sometimes refuse to follow benign instructions, such as declining to argue a political position or adopt a stated perso

safetyarxiv-cs-cl
3 Jun 2026
Safety

Infeasible optimization problems and the hierarchical augmented Lagrangian method in imitation learning

DGX agent

arXiv:2606.00730v1 Announce Type: new Abstract: Imitation learning (IL) is an effective approach to train complex robotics policies. Recent works have introduced hard constraints into imitation-learni

safetyarxiv-cs-ro
2 Jun 2026
Safety

Make Mechanistic Interpretability Auditable: A Call to Develop Guidelines via Continuous Collaborative Reviewing

DGX agent

arXiv:2606.00033v1 Announce Type: cross Abstract: While mechanistic interpretability (MI) has produced important insights into neural network internals, the field has yet to establish a standardized s

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safe2Drive: Evaluating Safe Driving Behaviors of E2E Autonomous Driving Models

DGX agent

arXiv:2606.00191v1 Announce Type: cross Abstract: Recent end-to-end (E2E) autonomous driving policies achieve high driving scores in closed-loop simulations. Yet it remains unclear whether these polic

safetyarxiv-cs-cv
2 Jun 2026
Safety

Tether-Aware Dynamic Collision Avoidance for USV-HROV Systems

DGX agent

arXiv:2606.01112v1 Announce Type: new Abstract: Heterogeneous marine robotic systems composed of an unmanned surface vehicle (USV) and a hybrid remotely operated vehicle (HROV) have shown great potent

safetyarxiv-cs-ro
2 Jun 2026
Safety

Stateful Online Monitoring Catches Distributed Agent Attacks

DGX agent

arXiv:2605.31593v1 Announce Type: cross Abstract: Language models can find thousands of severe software vulnerabilities, and agents are increasingly being misused for cyberattacks. To avoid detection,

safetyarxiv-cs-ai
1 Jun 2026
Safety

TARIC: Memory-Augmented Traversability-Aware Outdoor VLN under Interrupted Semantic Cues

DGX agent

arXiv:2605.31121v1 Announce Type: cross Abstract: Outdoor vision-language navigation (VLN) in long-range, open-world environments is frequently disrupted by semantic-cue interruptions, where informati

safetyarxiv-cs-ai
1 Jun 2026
Safety

Automating Low-Risk Code Review at Meta: RADAR, Risk Calibration, and Review Efficiency

DGX agent

arXiv:2605.30208v1 Announce Type: cross Abstract: AI-assisted coding tools have altered software production. At Meta, significant lines of code per human-landed diff grew by 105.9% year over year and

safetyarxiv-cs-ai
29 May 2026
← Previous
1…3132333435…257
Next →