AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

AI Native Games: A Survey and Roadmap

DGX agent

arXiv:2607.00527v1 Announce Type: new Abstract: Generative AI now enables games to produce dialogue, quests, characters, images, and worlds at runtime. Yet generation alone does not make a game AI-nat

safetyarxiv-cs-ai
2 Jul 2026
Safety

ECoSim: Data Efficient Fine-Tuning for Controllable Traffic Simulation

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.00545v1 Announce Type: new Abstract: Controllable traffic simulation is critical for testing autonomous driving systems, yet existing approaches often require retraining large generative mo

safetyarxiv-cs-cv
2 Jul 2026
Safety

Exploring the Semantic Gap in Agentic Data Systems: A Formative Study of Operationalization Failures in Analytical Workflows

DGX agent

arXiv:2607.00828v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate queries, invoke tools, and construct analytical workflows. Although recent advances hav

safetyarxiv-cs-ai
2 Jul 2026
Safety

From Holistic Evaluation to Structured Criteria: Rubrics Across the Evolving LLM Landscape

DGX agent

arXiv:2606.08625v2 Announce Type: replace Abstract: As Large Language Models (LLMs) advance toward open-ended autonomous agents, the mechanisms used to evaluate and guide their behavior must evolve ac

safetyarxiv-cs-cl
2 Jul 2026
Safety

Learning from Demonstration via Spatiotemporal Tubes for Unknown Euler-Lagrange Systems

DGX agent

arXiv:2607.00534v1 Announce Type: new Abstract: We present STT-LfD, a unified Learning from Demonstration (LfD) framework that integrates motion learning with control for unknown Euler-Lagrange system

safetyarxiv-cs-ro
2 Jul 2026
Safety

Learning Gait-Aware Quadruped Locomotion with Temporal Logic Specifications

DGX agent

arXiv:2607.00442v1 Announce Type: cross Abstract: Reinforcement learning (RL) for quadruped locomotion commonly depends on fixed, hand-crafted, and Markovian reward functions that limit both interpret

safetyarxiv-cs-ai
2 Jul 2026
Safety

Manifold-constrained Hamilton-Jacobi Reachability Learning for Decentralized Multi-Agent Motion Planning

DGX agent

arXiv:2511.03591v2 Announce Type: replace Abstract: Safe multi-agent motion planning (MAMP) under task-induced constraints is a critical challenge in robotics. Many real-world scenarios require robots

safetyarxiv-cs-ro
2 Jul 2026
Safety

Mnemosyne: Agentic Transaction Processing for Validating and Repairing AI-generated Workflows

DGX agent

arXiv:2607.00269v1 Announce Type: new Abstract: LLMs, solvers, and agent teams increasingly generate workflow actions, repairs, and plans, but a generated action may be syntactically valid yet stale,

safetyarxiv-cs-ai
2 Jul 2026
Safety

Reasoning Up the Instruction Ladder for Controllable Language Models

DGX agent

arXiv:2511.04694v5 Announce Type: replace-cross Abstract: As large language model (LLM) based systems take on high-stakes roles in real-world decision-making, they must reconcile competing instruction

safetyarxiv-cs-ai
2 Jul 2026
Safety

Surfacing Variations to Calibrate Perceived Reliability of MLLM-generated Image Descriptions

DGX agent

arXiv:2507.15692v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) provide new opportunities for blind and low vision (BLV) people to access visual information in their daily l

safetyarxiv-cs-cl
2 Jul 2026
Safety

Zero-Shot Distracted Driver Detection via Vision Language Models with Double Decoupling

DGX agent

arXiv:2601.08467v2 Announce Type: replace Abstract: Distracted driving is a major cause of traffic collisions, calling for robust and scalable detection methods. Vision-language models (VLMs) enable s

safetyarxiv-cs-cv
2 Jul 2026
Safety

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

DGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

safetyarxiv-cs-ai
1 Jul 2026
Safety

A Scalable Whole-body Motion Transfer via Implicit Kinodynamic Motion Retargeting

DGX agent

arXiv:2509.15443v2 Announce Type: replace-cross Abstract: Human-to-humanoid imitation learning presents a promising pathway to address the severe data scarcity bottleneck in robotics by utilizing abun

safetyarxiv-cs-ai
1 Jul 2026
Safety

A Tutorial on Autonomous Fault-Tolerant Control Using Knowledge-Grounded LLM Agents

DGX agent

arXiv:2606.31635v1 Announce Type: cross Abstract: Fault recovery in process plants still relies heavily on plant operators, especially when faults fall outside predefined supervisory logic. Operators

safetyarxiv-cs-ai
1 Jul 2026
Safety

AI for Quality Assurance in the Operating Room

DGX agent

arXiv:2606.30657v1 Announce Type: cross Abstract: Surgical outcomes depend not only on patient factors and postoperative care but are also strongly influenced by the quality of the operation itself. Y

safetyarxiv-cs-ai
1 Jul 2026
Safety

Certified Speculative Execution for Untrusted AI Agents

DGX agent

arXiv:2606.31023v1 Announce Type: cross Abstract: Hard-constrained sequential decision systems have no certified way to spend the test-time compute of modern AI: executing the multi-step drafts of a l

safetyarxiv-cs-lg
1 Jul 2026
Safety

Freeform Preference Learning for Robotic Manipulation

DGX agent

arXiv:2606.32027v1 Announce Type: cross Abstract: Reward design remains a central bottleneck for autonomous robot policy improvement, especially in long-horizon manipulation tasks where sparse success

safetyarxiv-cs-ai
1 Jul 2026
Safety

GRAPE: Graph-Augmented Prototype Explanations for Interactive Medical Image Diagnosis

DGX agent

arXiv:2606.30901v1 Announce Type: new Abstract: Prototype-based medical image classifiers present three clinical limitations: they treat findings as independent, silently amplify unsafe physician feed

safetyarxiv-cs-cv
1 Jul 2026
Safety

Online Generation of Collision-Free Trajectories in Dynamic Environments

DGX agent

arXiv:2603.00759v2 Announce Type: replace Abstract: In this paper, we present an online method for converting an arbitrary geometric path, represented by a sequence of states, and generated by any pla

safetyarxiv-cs-ro
1 Jul 2026
Safety

Relational and Sequential Conformal Inference for Energy Time Series over Graphs via Foundation Models

DGX agent

arXiv:2606.31804v1 Announce Type: new Abstract: Accurate energy demand forecasting is essential for the reliable operation and planning of modern sustainable energy systems. Spatial-temporal graph neu

safetyarxiv-cs-lg
1 Jul 2026
Safety

Toxicity Assessment in Preclinical Histopathology via Class-Aware Mahalanobis Distance for Known and Novel Anomalies

DGX agent

arXiv:2602.02124v2 Announce Type: replace-cross Abstract: Drug-induced toxicity is a leading cause of preclinical and early-clinical failure, making early detection critical. Histopathology is the gol

safetyarxiv-cs-ai
1 Jul 2026
Safety

Training Therapeutic Judges and Multi-Agent Systems for Human-Aligned Mental Health Support

DGX agent

arXiv:2606.30887v1 Announce Type: cross Abstract: Large language models show promise for mental health support, yet therapeutic quality improves only when evaluation functions as an actionable control

safetyarxiv-cs-ai
1 Jul 2026
Safety

AERMANI-VLM: Structured Prompting and Reasoning for Aerial Manipulation with Vision Language Models

DGX agent

arXiv:2511.01472v2 Announce Type: replace Abstract: The rapid progress of vision--language models (VLMs) has sparked growing interest in robotic control, where natural language can express the operati

safetyarxiv-cs-ro
30 Jun 2026
Safety

Budgeted Act-or-Defer Multi-Agent LLM Deliberation with Local Reliability Bounds

DGX agent

arXiv:2606.29654v1 Announce Type: new Abstract: Multi-agent deliberation among LLMs can improve reasoning, but deployment requires deciding when the current answer is reliable enough to act on and whe

safetyarxiv-cs-ai
30 Jun 2026
Safety

CaresAI at CT-DEB26: Detecting Dosing Errors In Clinical Trials Using Domain-Specific Transformer Embeddings and Classification Models

DGX agent

arXiv:2606.30236v1 Announce Type: new Abstract: Medication errors, particularly dosing errors in clinical trials (CT), can lead to patient harm, adverse drug events and worse patient outcomes. Dosing

safetyarxiv-cs-cl
30 Jun 2026
Safety

Characterizing Large Language Model Agentic Workflows: A Study on N8n Ecosystem

DGX agent

arXiv:2606.29116v1 Announce Type: new Abstract: Large Language Models (LLMs) are rapidly being adopted in low-code and no-code automation platforms, where non-expert users design workflows that combin

safetyarxiv-cs-ai
30 Jun 2026
Safety

Concept Removal Guidance: Evidence-Calibrated Negative Guidance for Safe Diffusion Sampling

DGX agent

arXiv:2606.29801v1 Announce Type: new Abstract: Text-to-image diffusion models remain vulnerable to adversarial prompts that elicit disallowed content, motivating reliable inference-time controls. A p

safetyarxiv-cs-cv
30 Jun 2026
Safety

Entity Binding Failures in Tool-Augmented Agents

DGX agent

arXiv:2606.30531v1 Announce Type: new Abstract: Tool-augmented language-model agents are often evaluated by whether they select the correct tool, produce valid API arguments, and complete the requeste

safetyarxiv-cs-ai
30 Jun 2026
Safety

Expert-guided Clinical Text Augmentation via Query-Based Model Collaboration

DGX agent

arXiv:2509.21530v2 Announce Type: replace Abstract: Data augmentation is a widely used strategy to improve model robustness and generalization by enriching training datasets with synthetic examples. W

safetyarxiv-cs-lg
30 Jun 2026
Safety

ExploreVLA: Dense World Modeling and Exploration for End-to-End Autonomous Driving

DGX agent

arXiv:2604.02714v2 Announce Type: replace Abstract: End-to-end autonomous driving models based on Vision-Language-Action (VLA) architectures have shown promising results by learning driving policies t

safetyarxiv-cs-cv
30 Jun 2026
Safety

Flying to Image-Specified Objects: 3D Quadrotor Navigation via Cross-Graph Memory and Viewpoint Planning

DGX agent

arXiv:2606.29917v1 Announce Type: new Abstract: Instance-Specific Image-Goal Navigation (InstanceImageNav) requires a robot to navigate toward the exact object instance depicted in a query image. Exte

safetyarxiv-cs-ro
30 Jun 2026
Safety

IHDec: Divergence-Steered Contrastive Decoding for Securing Multi-Turn Instruction Hierarchies

DGX agent

arXiv:2606.29960v1 Announce Type: new Abstract: Large Language Models (LLMs) often fail to maintain instruction hierarchies (IH) when processing multi-source inputs with varying role-level priorities,

safetyarxiv-cs-cl
30 Jun 2026
Safety

Langshaw: Declarative Interaction Protocols Based on Sayso and Conflict

DGX agent

arXiv:2606.29601v1 Announce Type: cross Abstract: Current languages for specifying multiagent protocols either over-constrain protocol enactments or complicate capturing their meanings. We propose Lan

safetyarxiv-cs-ai
30 Jun 2026
Safety

LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing

DGX agent

arXiv:2507.07056v2 Announce Type: replace-cross Abstract: The proliferation of Low-Rank Adaptation (LoRA) models has democratized personalized text-to-image generation, enabling users to share lightwe

safetyarxiv-cs-lg
30 Jun 2026
Safety

Mechanistically Eliciting Latent Behaviors in Language Models

DGX agent

arXiv:2606.29604v1 Announce Type: cross Abstract: We aim to discover diverse, generalizable perturbations of LLM internals that can surface hidden behavioral modes. Such perturbations could help resha

safetyarxiv-cs-ai
30 Jun 2026
Safety

MOAR Planner: Multi-Objective and Adaptive Risk-Aware Path Planning for Infrastructure Inspection with a UAV

DGX agent

arXiv:2606.30575v1 Announce Type: new Abstract: The problem of autonomous navigation for UAV inspection remains challenging as it requires effectively navigating in close proximity to obstacles, while

safetyarxiv-cs-ro
30 Jun 2026
Safety

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

DGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

safetyarxiv-cs-ai
30 Jun 2026
Safety

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

DGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

safetyarxiv-cs-ai
30 Jun 2026
Safety

OWMDrive: Causality-Aware End-to-End Autonomous Driving via 4D Occupancy World Model

DGX agent

arXiv:2606.30421v1 Announce Type: new Abstract: Autonomous driving systems are steadily moving toward end-to-end paradigms to mitigate the limited adaptability of rule-based pipelines in complex traff

safetyarxiv-cs-cv
30 Jun 2026
Safety

PL-LIT: A LiDAR-Inertial-Thermal SLAM Using Point-Line Features and Thermographic Mapping

DGX agent

arXiv:2606.29259v1 Announce Type: new Abstract: Thermal imaging is resilient to adverse conditions, such as intense illumination, low-light operation, and fog, and can therefore mitigate odometry degr

safetyarxiv-cs-ro
30 Jun 2026
Safety

Pose-Based Fall Detection System: Efficient Monitoring on Standard CPUs

DGX agent

arXiv:2503.19501v2 Announce Type: replace-cross Abstract: Falls among elderly residents in assisted living homes pose significant health risks, often leading to injuries and a decreased quality of lif

safetyarxiv-cs-ai
30 Jun 2026
Safety

Propagation of~Interval Belief Structures and~Imprecise Copulas for~Neural Network Verification

DGX agent

arXiv:2606.30105v1 Announce Type: new Abstract: Quantitative verification of neural networks requires reasoning about probabilities under substantial uncertainty in both input distributions and their

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

SafePyramid: A Hierarchical Benchmark for In-context Policy Guardrailing

DGX agent

arXiv:2606.29887v1 Announce Type: new Abstract: In real-world applications, guardrails are often expected to identify unsafe user-model interactions according to application-specific safety policies,

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Self-Organized Conformal Prediction: Reducing Regional Coverage Gaps with Unsupervised Group Discovery

DGX agent

arXiv:2606.29403v1 Announce Type: cross Abstract: Conformal prediction guarantees marginal coverage, but pooled calibration averages over heterogeneous regions and can mask regional undercoverage in s

safetyarxiv-cs-ai
30 Jun 2026
Safety

Test-Time Detoxification without Training or Learning Anything

DGX agent

arXiv:2602.02498v2 Announce Type: replace-cross Abstract: Large language models can produce toxic or inappropriate text even for benign inputs, creating risks when deployed at scale. Detoxification is

safetyarxiv-cs-ai
30 Jun 2026
Safety

The Alignment Auditor: A Bayesian Framework for Verifying and Refining LLM Objectives

DGX agent

arXiv:2510.06096v3 Announce Type: replace-cross Abstract: The objectives that Large Language Models (LLMs) implicitly optimize remain dangerously opaque, making trustworthy alignment and auditing a gr

safetyarxiv-cs-cl
30 Jun 2026
Safety

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

DGX agent

arXiv:2606.28425v1 Announce Type: cross Abstract: Increasingly autonomous agentic AI systems pose novel multi-agent risks, such as secret collusion via covert communication channels. The natural defen

safetyarxiv-cs-ai
30 Jun 2026
Safety

TrajRS: Towards Certified Robustness in Pedestrian Trajectory Prediction

DGX agent

arXiv:2606.28716v1 Announce Type: new Abstract: The robustness of trajectory prediction models is crucial for developing safe autonomous driving systems. Adversarial attacks on trajectory prediction c

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…4243444546…257
Next →