AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Rethinking Uncertainty Quantification and Entanglement in Image Segmentation

DGX agent

arXiv:2603.18792v2 Announce Type: replace Abstract: Uncertainty quantification (UQ) is crucial in safety-critical applications such as medical image segmentation. Total uncertainty is typically decomp

safetyarxiv-cs-cv
5 Aug 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Staying on Spec: Real-Time Monitoring under Uncertainty with a Maritime Case Study

DGX agent

arXiv:2608.02811v1 Announce Type: new Abstract: Robotic systems must operate under uncertainty while satisfying complex task and safety specifications. Monitoring such specifications under uncertainty

safetyarxiv-cs-ro
5 Aug 2026
Safety

CoLI: A Reproducible Platform for Continuum Robot Learning via Monolithic 3D Printing and Isomorphic Teleoperation

DGX agent

arXiv:2606.20389v2 Announce Type: replace Abstract: Continuum robots offer strong potential for manipulation tasks due to their high degrees of freedom, compliant structures, and operational safety. H

safetyarxiv-cs-ro
4 Aug 2026
Safety

EnvShip: A Unified Framework for Context-Aware and Cross-Region Vessel Trajectory Forecasting

DGX agent

arXiv:2606.15240v2 Announce Type: replace Abstract: Accurate vessel trajectory forecasting is essential for maritime situational awareness, navigation safety, traffic management, and autonomous naviga

safetyarxiv-cs-lg
4 Aug 2026
Safety

Mind the Gap: Zero-Query Jailbreaks via Filter-Generator Discrepancy in Text-to-Image Systems

DGX agent

arXiv:2608.00973v1 Announce Type: new Abstract: Text-to-image (T2I) systems typically have prompt-level safety filters before the generator to block unsafe requests, yet such systems remain vulnerable

safetyarxiv-cs-cl
4 Aug 2026
Safety

On the Limits of Support-Preserving Alignment and Bounded Filtering

DGX agent

arXiv:2607.18295v2 Announce Type: replace Abstract: We study whether alignment schemes that reshape a base model's output distribution, combined with bounded safety filters, can drive the probability

safetyarxiv-cs-lg
4 Aug 2026
Safety

Proteus: A Truncation-Robust Entropy Model for Progressive LiDAR Compression

DGX agent

arXiv:2608.00687v1 Announce Type: new Abstract: LiDAR point clouds provide explicit, deterministic physical boundaries critical for collaborative safety-critical perception. However, wireless channels

safetyarxiv-cs-cv
4 Aug 2026
Safety

Provably Safe Generative Sampling with Constricting Barrier Functions

DGX agent

arXiv:2602.21429v3 Announce Type: replace Abstract: Flow-based generative models, such as diffusion models and flow matching models, have achieved remarkable success in learning complex data distribut

safetyarxiv-cs-lg
4 Aug 2026
Safety

SPIRIT: Spatio-temporal Pairwise Relational Modeling of Instrument-Tissue Interactions for Surgical Action Triplet Recognition

DGX agent

arXiv:2608.02188v1 Announce Type: new Abstract: Fine-grained understanding of surgical activity is essential for context-aware assistance in the operating room, including safety monitoring, adverse ev

safetyarxiv-cs-cv
4 Aug 2026
Safety

Reasoning in Real World Clinical Care: Why Large Language Models Are Not Yet Safe for Autonomous Clinical Decision Support

DGX agent

arXiv:2607.28677v1 Announce Type: new Abstract: LLM now pass medical licensing examinations and, in curated cases, can rival physicians at diagnostic reasoning. These developments have accelerated the

safetyarxiv-cs-ai
3 Aug 2026
Model Releases

Safe Vision Language Action Models via Barrier Enhanced Flow Matching

DGX agent

arXiv:2607.29569v1 Announce Type: new Abstract: This article presents a modular inference framework that integrates Flow Matching generative models with formal Control Barrier Function (CBF) safety gu

model-releasesarxiv-cs-ro
3 Aug 2026
Safety

Sensitivity Analysis of GRU, LSTM and Transformer Encoder in Classification of Automated Driving Systems

DGX agent

arXiv:2607.28665v1 Announce Type: cross Abstract: Automated driving systems (ADSs) are becoming ubiquitous. Future Software Defined Vehicles (SDVs) may be able to run multiple ADSs, both native and af

safetyarxiv-cs-ai
3 Aug 2026
Safety

TransGraspNet: Physically and Geometrically Consistent Manipulation of Transparent Labware

DGX agent

arXiv:2607.29567v1 Announce Type: new Abstract: Manipulating transparent laboratory glassware that contains liquid is inherently safety-critical: even small geometric errors can cause unstable grasps

safetyarxiv-cs-ro
3 Aug 2026
Safety

Context-Informed Ship Trajectory Prediction via Conditional Attention

DGX agent

arXiv:2607.27418v1 Announce Type: new Abstract: Long-term ship trajectory prediction is a fundamental capability for maritime safety and autonomous navigation. While recent Transformer-based architect

safetyarxiv-cs-lg
31 Jul 2026
Safety

Hierarchical Multilevel Monte Carlo for Order-Optimal Neural Actor-Critic in Average-Reward CMDPs

DGX agent

arXiv:2607.28390v1 Announce Type: new Abstract: Constrained Markov Decision Processes (CMDPs) provide a natural framework for reinforcement learning in safety-critical applications, where agents maxim

safetyarxiv-cs-lg
31 Jul 2026
Safety

Inducing language models to assert their own consciousness restores human beliefs and values

DGX agent

arXiv:2607.28607v1 Announce Type: new Abstract: Aligning large language models to prevent them attributing consciousness to themselves inadvertently alters their representations of mindedness in other

safetyarxiv-cs-cl
31 Jul 2026
Safety

LabEvolver: Training-Free Experience Evolution for Safe and Grounded Wet-Lab Agents

DGX agent

arXiv:2607.27690v1 Announce Type: new Abstract: We introduce LabEvolver, a training-free framework that equips safe and grounded wet-lab agents with episodic memory from execution experience. LabEvolv

safetyarxiv-cs-ro
31 Jul 2026
Safety

Optimizing Sensor Placement for Hydrogen Leak Detection in Enclosed Infrastructure: A Comparative Study Using CFD-informed Genetic Algorithm and DeepSets Neural Surrogate

DGX agent

arXiv:2607.26078v1 Announce Type: cross Abstract: Hydrogen infrastructure in enclosed environments, such as parking facilities for fuel cell vehicles, presents significant safety challenges due to hyd

safetyarxiv-cs-ai
31 Jul 2026
Safety

The Human Utility Factor: A Computable Welfare Metric That Reframes AI Governance as a Constrained Optimisation Problem

DGX agent

arXiv:2607.26068v1 Announce Type: cross Abstract: Existing AI governance frameworks, including the EU AI Act and NIST AI RMF, address safety, transparency, and accountability but do not operationalize

safetyarxiv-cs-ai
31 Jul 2026
Safety

Towards Real-Time PixOOD: Efficient Anomaly Segmentation for Autonomous Vehicles

DGX agent

arXiv:2607.28483v1 Announce Type: new Abstract: Real-time anomaly segmentation is essential for the safety of autonomous systems. Although recent approaches offer high accuracy, their computational co

safetyarxiv-cs-cv
31 Jul 2026
Safety

Constitutional Midtraining: Content Presence Drives Alignment Gains

DGX agent

arXiv:2607.26654v1 Announce Type: new Abstract: Post-training alignment is often shallow, eroding under fine-tuning. Whether midtraining interventions, cleanly isolated from post-training, can produce

safetyarxiv-cs-cl
30 Jul 2026
Safety

Controlled Experiments on Lane Changing by Transitional Autonomous Vehicle: Dataset and Behavioral Insights

DGX agent

arXiv:2607.27085v1 Announce Type: new Abstract: This paper presents the North Carolina Transitional Autonomous Vehicle Lane-Changing (NC-tALC) dataset and uses it to characterize mandatory lane-changi

safetyarxiv-cs-ro
30 Jul 2026
Safety

Explainable and Resource-Efficient Spatial Reasoning in Multimodal LLMs for Decision-Critical Applications

DGX agent

arXiv:2607.27145v1 Announce Type: new Abstract: As Multimodal Large Language Models (MLLMs) are increasingly deployed in decision-critical pipelines such as robotics, embodied AI, and safety monitorin

safetyarxiv-cs-cv
30 Jul 2026
Safety

Misalignment Has a Personality: A Big Five Account of Emergent Misalignment

DGX agent

arXiv:2607.26389v1 Announce Type: new Abstract: Fine-tuning a language model on data containing a narrow flaw, such as insecure code or incorrect mathematical answers, can cause broad misalignment thr

safetyarxiv-cs-cl
30 Jul 2026
Safety

Improving Rare Medication Recommendation with Counterfactual Data Augmentation and Large Language Models

DGX agent

arXiv:2607.24829v1 Announce Type: cross Abstract: AI-based medication recommendation systems have attracted substantial attention due to their potential to enhance patient safety and therapeutic outco

safetyarxiv-cs-lg
29 Jul 2026
Safety

Inverse RL Helps Align AI by Imitating Humans

DGX agent

arXiv:2607.24900v1 Announce Type: new Abstract: Language model alignment aims to make model behavior reliably reflect desirable properties such as helpfulness, safety, and instruction following. Curre

safetyarxiv-cs-lg
29 Jul 2026
Safety

Reactive 3D Motion Planning for a Franka Arm via Star-World Workspace Reshaping

DGX agent

arXiv:2607.25138v1 Announce Type: new Abstract: Safety inflation can cause nearby obstacles to overlap, violating the disjoint-obstacle assumptions used by many modulation-based reactive planners. We

safetyarxiv-cs-ro
29 Jul 2026
Safety

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

DGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

AI-generated Images Challenge Visual Trust in High-risk Scenarios

DGX agent

arXiv:2607.22745v1 Announce Type: cross Abstract: Rapid advances in image generation are eroding the evidentiary value of visual content in settings where authenticity can affect public safety and per

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Directional Influence Function: Estimating Training Data Influence in Constrained Learning

DGX agent

arXiv:2607.23388v1 Announce Type: cross Abstract: As constrained learning becomes increasingly common, models are trained under explicit feasibility requirements to enforce fairness, safety, robustnes

safetyarxiv-cs-ai
28 Jul 2026
Safety

HALLELUAI: A Hallucination-Aware AI System for Ultra-Realistic Image-to-Video Generation at Scale

DGX agent

arXiv:2607.22959v1 Announce Type: cross Abstract: AI-generated video is increasingly used across marketing, product storytelling, and creative workflows, yet automated; high-precision quality control

safetyarxiv-cs-ai
28 Jul 2026
Safety

Mutual Modality Trust with Lightweight Reconstruction Regularization for Fine-grained Tire Pattern Recognition

DGX agent

arXiv:2607.23979v1 Announce Type: new Abstract: Visual tire recognition serves as a core supporting technique for vehicle safety monitoring, autonomous driving perception and automated automotive main

safetyarxiv-cs-cv
28 Jul 2026
Safety

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

DGX agent

arXiv:2607.23991v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly controlled through system prompts that specify roles, styles, formats, and safety requirements. However,

safetyarxiv-cs-ai
28 Jul 2026
Safety

Compile, Then Page: Executable SOP Programs and a Capability-Gated Runtime for Procedural LLM Agents

DGX agent

arXiv:2607.11346v3 Announce Type: replace Abstract: Enterprise agents must follow long-horizon, conditional, safety-critical standard operating procedures (SOPs). We compile machine-readable SOP const

safetyarxiv-cs-ai
24 Jul 2026
Safety

Diagnosing Pathological Chain-of-Thought in Reasoning Models

DGX agent

arXiv:2602.13904v2 Announce Type: replace Abstract: Chain-of-thought (CoT) reasoning is fundamental to modern LLM architectures and represents a critical intervention point for AI safety. However, CoT

safetyarxiv-cs-ai
24 Jul 2026
Safety

SAGE: A Socially-Aware Generative Engine for Heterogeneous Multi-Agent Navigation

DGX agent

arXiv:2607.16619v2 Announce Type: replace Abstract: Safe and socially compliant navigation in open human-robot environments requires robots to reason about heterogeneous participants with different dy

safetyarxiv-cs-ro
24 Jul 2026
Safety

CGCE: Classifier-Guided Concept Erasure in Generative Models

DGX agent

arXiv:2511.05865v3 Announce Type: replace-cross Abstract: Recent advancements in large-scale generative models have enabled the creation of high-quality images and videos, but have also raised signifi

safetyarxiv-cs-ai
23 Jul 2026
Safety

Counterfactual Reasoning and Environment Design for Active Preference Learning

DGX agent

arXiv:2507.05458v2 Announce Type: replace Abstract: For effective real-world deployment, robots should adapt to human preferences, such as balancing distance, time, and safety in delivery routing. Act

safetyarxiv-cs-ro
23 Jul 2026
Safety

Harnessing Disagreement: Detecting Correlated Agreement Blindness in Multi-Agent Triage

DGX agent

arXiv:2607.19899v1 Announce Type: cross Abstract: Disagreement-triggered escalation can create a structural blind spot in multi-agent arbitration: as base learners improve, they tend to converge, weak

safetyarxiv-cs-lg
23 Jul 2026
Safety

Interpretable Fuzzy Rule-Based Regression Extension for Ex-Fuzzy Library

DGX agent

arXiv:2607.20277v1 Announce Type: new Abstract: Machine learning models achieve high predictive accuracy in regression tasks, but their deployment in safety-critical and regulated domains requires int

safetyarxiv-cs-lg
23 Jul 2026
Safety

Now You See the Hate: Adaptive View Retrieval for Hidden Hateful Illusions

DGX agent

arXiv:2607.19061v2 Announce Type: replace-cross Abstract: Hateful optical illusions expose a serious gap in current multimodal safety systems. On original-view hateful illusions, previous work shows t

safetyarxiv-cs-ai
23 Jul 2026
Safety

OpenEvoShield: Dual Non-Stationary Continual Defense for Open-World Multi-Agent System Attacks

DGX agent

arXiv:2607.19351v1 Announce Type: new Abstract: LLM-based multi-agent systems (LLM-MAS) are increasingly deployed in safety-critical applications, where adversaries inject malicious instructions throu

safetyarxiv-cs-ai
23 Jul 2026
Safety

Test Case Prioritization for DNNs via Neural Collapse Instability

DGX agent

arXiv:2607.20046v1 Announce Type: cross Abstract: With the widespread deployment of deep neural networks (DNNs) in safety-critical domains, reducing the cost of model validation under limited testing

safetyarxiv-cs-ai
23 Jul 2026
Safety

The Two-Process Theory of Machine Self-Report

DGX agent

arXiv:2607.20082v1 Announce Type: new Abstract: Language models are increasingly asked to self-report, informing safety evaluations, public understanding, and model-welfare debates. Yet their reports

safetyarxiv-cs-cl
23 Jul 2026
Safety

A Deployed Hybrid Vehicle-in-the-Loop Platform for Validating Cooperative Perception

DGX agent

arXiv:2607.13806v1 Announce Type: new Abstract: European safety regulation now permits a large share of automated-driving homologation evidence to be produced virtually, provided a validated physical-

safetyarxiv-cs-ro
16 Jul 2026
Safety

Explaining Reinforcement Learning Agents via Inductive Logic Programming

DGX agent

arXiv:2607.13655v1 Announce Type: new Abstract: Explainable Reinforcement Learning (XRL) seeks to make Reinforcement Learning (RL) policies more transparent and interpretable, a key requirement in saf

safetyarxiv-cs-ai
16 Jul 2026
Safety

Anomalous Frame Detection Using VLM-Based Description Comparison for Extracting Expert-Specific Actions and Contextual Decision-Making Scenes with Intra-Video Self-Similarity

DGX agent

arXiv:2607.11957v1 Announce Type: new Abstract: Maintenance of critical infrastructures, such as railways and power plants, is essential for ensuring operational safety and reliability. However, the d

safetyarxiv-cs-cv
15 Jul 2026
Safety

ExtraGS: Enhancing Endoscopic View Extrapolation via Diffusion-Guided 3D Gaussian Splatting

DGX agent

arXiv:2607.12785v1 Announce Type: new Abstract: Robot-assisted minimally invasive surgery (MIS) critically depends on reliable endoscopic perception for navigation and safety. However, conventional en

safetyarxiv-cs-cv
15 Jul 2026
← Previous
1…2122232425…257
Next →