AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
11 Aug 2026

Agentic Visual Reasoning in Whole-Slide Pathology Images via Active Perception

SafetyDGX agent

arXiv:2608.08648v1 Announce Type: new Abstract: Whole-slide visual reasoning requires identifying sparse diagnostic evidence in gigapixel pathology slides and integrating observations across spatial s

An Agentic Generative Large Language Model for Treatment Planning of Colorectal Cancer

SafetyDGX agent

arXiv:2608.09142v1 Announce Type: new Abstract: Treatment planning in precision oncology requires synthesizing heterogeneous patient information with rapidly evolving clinical guidelines to ensure gui

An Explainable GNN Framework for Component-Level Anomaly Diagnosis

SafetyDGX agent

arXiv:2608.09246v1 Announce Type: new Abstract: Industrial processes are complex systems composed of multiple interacting sensors that generate multivariate time series (MTS). Detecting anomalies in s


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ArchAgent v2: A Case Study with the Data Prefetching Championship

SafetyDGX agent

arXiv:2608.09874v1 Announce Type: new Abstract: Agentic artificial intelligence has shown great promise in automating algorithm design, but scaling similar techniques to computer microarchitecture dis

Artificial Leviathan: Exploring Social Evolution of LLM Agents Through the Lens of Hobbesian Social Contract Theory

SafetyDGX agent

arXiv:2406.14373v3 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) and advancements in Artificial Intelligence (AI) offer an opportunity for computational social science

Auditing Instruction-Trajectory Mismatches in Multimodal Robot Demonstrations

SafetyDGX agent

arXiv:2608.07895v1 Announce Type: cross Abstract: Robot demonstration datasets used to train vision-language-action policies can contain a subtle but harmful failure mode: trajectories that are behavi

Autonomous Driving with Priority-Ordered STL Specifications Under Multimodal Uncertainty

SafetyDGX agent

arXiv:2606.20336v2 Announce Type: replace Abstract: Autonomous vehicles must plan trajectories that satisfy multiple requirements, such as safety, traffic-rule compliance, and passenger comfort. Howev

Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents

SafetyDGX agent

arXiv:2510.04465v3 Announce Type: replace-cross Abstract: LLM agents require personal information for personalization in order to effectively act on users' behalf, but this raises privacy concerns tha

Beyond Aggregate Calibration: Decomposing Income-Conditional Recall Disparities in Automated Credit Default Prediction

SafetyDGX agent

arXiv:2608.08202v1 Announce Type: new Abstract: Data-centric curation pipelines frequently rely on model confidence scores to flag and filter noisy or mislabeled training instances. Evaluating this fi

Beyond Binary: Continuous State Optimization with Graph-Structured Objectives

SafetyDGX agent

arXiv:2608.09366v1 Announce Type: new Abstract: Large-scale learning systems often face the challenge of balancing multiple, potentially competing objectives, such as fairness, accuracy, and latency.

Beyond cognacy

SafetyDGX agent

arXiv:2507.03005v3 Announce Type: replace Abstract: Computational phylogenetics has become an established tool in historical linguistics, with many language families now analyzed using likelihood-base

Beyond 'I Can't Help With That': How Child Safety Experts Evaluate AI Chatbot Safety

SafetyDGX agent

arXiv:2608.07902v1 Announce Type: cross Abstract: Youth increasingly turn to AI chatbots for social and emotional support, raising concerns about how these systems respond, especially in high-stakes s

Beyond Solvability: Task Learnability as a Static Prior for LLM RL Post-Training

SafetyDGX agent

arXiv:2608.09217v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central post-training paradigm for eliciting reasoning capabilities in large language models, yet uniform tas

Causal Falsification of Digital Twins

SafetyDGX agent

arXiv:2301.07210v5 Announce Type: replace-cross Abstract: Digital twins are simulation-based models designed to predict how a real-world process will evolve in response to interventions. This modellin

CDGC-Net: 3D Medical Image Segmentation with Cooperative Dual-Scale Self-Attention and Grouped Channel Modeling

SafetyDGX agent

arXiv:2608.08575v1 Announce Type: cross Abstract: Accurate 3D medical image segmentation requires the integration of long-range anatomical context with fine boundary detail. Existing methods often mod

CIFA: Contextual-Intersectional Fairness Auditing for Hidden Subgroup Discovery in Face Analysis

SafetyDGX agent

arXiv:2608.09669v1 Announce Type: new Abstract: Fairness evaluation in computer vision commonly relies on aggregate accuracy and demographic subgroup analysis. However, visual models are also sensitiv

City Sentinel: A Unified AI-Based Smart Surveillance Framework for Real-Time Multi-Threat Detection Using Deep Learning

SafetyDGX agent

arXiv:2608.08887v1 Announce Type: new Abstract: Rapid urbanization has increased the need for surveillance systems that can monitor multiple public safety risks at the same time. Traditional systems o

CLAM: Causal Spatial Disaggregation to Infer Local Effects From Coarse Data

SafetyDGX agent

arXiv:2608.08064v1 Announce Type: new Abstract: Learning fine-grained spatial patterns from coarse-resolution data is challenging, especially in causal settings where high-resolution effects must be i

Coarse-to-Fine Registration of Jawbone CT and Intraoral Scan Data Using GeDi and ICP with Pseudo-IOS Ground Truth

SafetyDGX agent

arXiv:2608.07564v1 Announce Type: cross Abstract: In digital dentistry and oral surgery, the registration of jawbone CT and intraoral scanner (IOS) data is essential for integrating internal bone stru

Concept-Guided Spatial Regularization for World Models in Atari Pong

SafetyDGX agent

arXiv:2607.15142v2 Announce Type: replace Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understu

Confusion-Geometry Rebalancing for Long-Tailed Adversarial Training

SafetyDGX agent

arXiv:2608.09688v1 Announce Type: cross Abstract: Adversarial training under long tailed distributions suffers from a dual imbalance: the class imbalance skews the training objective toward head class

Context Is Not Authority: Structured Runtime Governance for Financial Market Agents

SafetyDGX agent

arXiv:2608.09025v1 Announce Type: new Abstract: Financial agents can turn correct context into an unauthorized effect: a customer-facing commitment, trade, or deployed policy. We present SAGE-Fin, a f

Contextual Value Alignment via Multilayer Combinatorial Fusion

SafetyDGX agent

arXiv:2608.07642v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF

Control-Oriented Scenario Tree Construction through Reinforcement Learning

SafetyDGX agent

arXiv:2608.09335v1 Announce Type: new Abstract: Multistage stochastic model predictive control (MPC) handles uncertainty by optimizing over a scenario tree, a finite branching approximation of future

Coordinated incentives in AI-generated misinformation governance

SafetyDGX agent

arXiv:2608.07070v1 Announce Type: cross Abstract: With the rapid diffusion of AI-generated content, AI-driven misinformation is becoming increasingly pervasive and difficult to govern, undermining inf

CoRCi: Cross-Reconstruction of Coherent Interests Modeling in Cross-Domain Sequential Recommendation

SafetyDGX agent

arXiv:2608.09580v1 Announce Type: new Abstract: Cross-Domain Sequential Recommendation (CDSR) aims to alleviate data sparsity by transferring dynamic user interests across related domains. A key chall

Correlation flow governs learning at criticality

SafetyDGX agent

arXiv:2608.08350v1 Announce Type: new Abstract: The initialisation of deep neural networks determines whether information and gradients can propagate across depth, yet a unified theory connecting thes

CPDA: Class-Conditional Path Distribution Alignment for Unsupervised Time-Series Domain Adaptation

SafetyDGX agent

arXiv:2608.09193v1 Announce Type: cross Abstract: Unsupervised time-series domain adaptation (DA) addresses the challenge of transferring a classifier from a labeled source domain to an unlabeled targ

Critic-Free Deep Reinforcement Learning for Maritime Coverage Path Planning on Irregular Hexagonal Grids

SafetyDGX agent

arXiv:2603.28385v2 Announce Type: replace-cross Abstract: Maritime surveillance missions, such as search and rescue and environmental monitoring, rely on the efficient allocation of sensing assets ove

Crowd-Sourced Geographies of Income: Using Google Maps Points of Interest as High-Frequency Proxies for Sub-Municipal Income Estimation in Sao Paulo, Brazil

SafetyDGX agent

arXiv:2608.07871v1 Announce Type: cross Abstract: Accurate, up-to-date income data at the sub-municipal scale is essential for social policy in middle-income countries, yet in Brazil it depends on a c

CUPA-T2*: Covariance-Aware Uncertainty Propagation and Alignment for T2* Mapping in Accelerated MRI

SafetyDGX agent

arXiv:2608.08693v1 Announce Type: new Abstract: Quantitative T2* maps have strong potential for biomarker discovery but are limited by long scan times, rendering them impractical in clinical settings.

CyberAGENTS: Structured Autonomy for Agentic Gamified Learning in Cybersecurity

SafetyDGX agent

arXiv:2608.07965v1 Announce Type: new Abstract: Gamification is especially effective in learning domains requiring active problem-solving and iterative skill-building, such as cybersecurity education.

DA-NBV: A Direction-Aware Next-Best-View Planner for Efficient 3D Reconstruction of Ships at Sea

SafetyDGX agent

arXiv:2608.08025v1 Announce Type: cross Abstract: Accurate 3D reconstruction of ships at sea is important for maritime supervision, damage assessment, and autonomous maritime operations. Although 3D r

Decoy Images Amplify Caption-Mediated Defenses Against Encoded Jailbreaks

SafetyDGX agent

arXiv:2608.01043v2 Announce Type: replace-cross Abstract: We report a counter-intuitive interaction between image inputs and existing black-box defenses on Vision--Language Models (VLMs): pairing an e

Demystifying Prediction Powered Inference

SafetyDGX agent

arXiv:2601.20819v2 Announce Type: replace-cross Abstract: Machine learning predictions are increasingly used to supplement incomplete or costly-to-measure outcomes in fields such as biomedical researc

Dense Point-to-Mask Optimization with Reinforced Point Selection for Crowd Instance Segmentation

SafetyDGX agent

arXiv:2604.01742v2 Announce Type: replace Abstract: Crowd instance segmentation is a crucial task with a wide range of applications, including surveillance and transportation. Currently, point labels

Deployable Per-Instance Multi-Layer Activation Steering for Large Language Models

SafetyDGX agent

arXiv:2608.08829v1 Announce Type: cross Abstract: Activation steering edits the behaviour of a frozen language model by adding a learned vector to its residual stream, and current practice fixes the i

Designing for Ethical AI: HCI Feature Considerations to Improve Fairness and User Experience in AutoML use for Human Resources

SafetyDGX agent

arXiv:2608.07477v1 Announce Type: cross Abstract: This thesis examines the fairness of Automated Machine Learning (AutoML) tools in human resource hiring systems through the combined lenses of regulat

Determinization in Structure Theories: A Unified Framework via Closure, Comparability, and Joint Admissibility

SafetyDGX agent

arXiv:2608.07476v1 Announce Type: new Abstract: We develop a formal framework for constructing canonical interpretations from plural structure theories. A structure theory is a triple T = ({Sigma}, A,

DH-VLM: Dual-Horizon Cooperative Latent Reasoning for Autonomous Driving

SafetyDGX agent

arXiv:2608.09333v1 Announce Type: new Abstract: Large-scale language models for autonomous driving enable enhanced global understanding and long-horizon planning. However, when deployed in isolated ve

DiSCo: Diffusion Sequence Copilots for Shared Autonomy

SafetyDGX agent

arXiv:2603.22787v2 Announce Type: replace-cross Abstract: Shared autonomy combines human user and AI copilot actions to control complex systems such as robotic arms. When a task is challenging, requir

Discovering and Causally Validating Emotion-Sensitive Neurons in Large Audio-Language Models

SafetyDGX agent

arXiv:2601.03115v2 Announce Type: replace Abstract: Emotion is a central dimension of spoken communication, yet, we still lack a mechanistic account of how modern large audio-language models (LALMs) e

Distill Skills into Weights, Not Prompts: Abstract Skills as Privileged Signals for On-Policy Self-Distillation

SafetyDGX agent

arXiv:2608.09826v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards yields no group-relative signal when rollout groups are uniformly correct or uniformly wrong, which acc

Distilling Vision-Language Models for Robust Traffic Sign Perception in Autonomous Vehicles

SafetyDGX agent

arXiv:2608.08815v1 Announce Type: new Abstract: Traffic sign recognition (TSR) models based on deep neural networks achieve strong clean-data performance but remain vulnerable to physically realizable

Distributed Optimization with Streaming Data: A Temporal Weighting Perspective

SafetyDGX agent

arXiv:2608.09565v1 Announce Type: cross Abstract: Optimization theory is a widely used tool for intelligent decision-making. While classical optimization deals with fixed, time-invariant objective fun

Distribution-Free Conformal Prediction for Steel Fatigue Strength: Marginal Validity Is Not Enough

SafetyDGX agent

arXiv:2608.07589v1 Announce Type: cross Abstract: Predicting fatigue failure in steel components experimentally is costly because it requires testing across multiple compositions and processing condit

DoGMA: A Central-Dogma-Guided Foundation Model for Multi-Omics Alignment and Multi-Task Learning in Oncology

SafetyDGX agent

arXiv:2608.08148v1 Announce Type: cross Abstract: Attention mechanisms have been widely utilized in modern deep learning, and many existing multi-omics models inherit their conventional use to allow u

DreOPD: Degraded-Reference Extrapolative On-Policy Distillation for Flow-matching Models

SafetyDGX agent

arXiv:2608.09233v1 Announce Type: cross Abstract: Flow-matching models are now a mainstream method to image generation, but its adaptation to diverse downstream scenarios typically relies on post-trai

DSLE: A Learning Environment for Dark Souls Boss Encounters

SafetyDGX agent

arXiv:2608.09902v1 Announce Type: new Abstract: We introduce the Dark Souls Learning Environment (DSLE), a containerized platform that presents all 22 boss encounters of Dark Souls: Remastered as game

Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs

SafetyDGX agent

arXiv:2608.09542v1 Announce Type: cross Abstract: Large reasoning models (LRMs) achieve remarkable success on complex tasks but remain vulnerable to harmful prompts that induce unsafe outputs. Recent

Dynamic Distribution-Aware Uncertainty Tracking in Vision-Language Representation Learning

SafetyDGX agent

arXiv:2608.09011v1 Announce Type: new Abstract: Uncertainty Quantification (UQ) aims to measure the reliability of model predictions, serving as a critical safeguard for deploying Vision-Language Mode

EFFEKT: Efficient Federated Knowledge Transfer to Foundation Models

SafetyDGX agent

arXiv:2608.08138v1 Announce Type: new Abstract: Recent data protection laws have accelerated the adoption of Federated Learning (FL) for privacy-preserving decentralized training. Nevertheless, increa

Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition

SafetyDGX agent

arXiv:2608.09762v1 Announce Type: new Abstract: Real-world online reinforcement learning (RL) provides a promising approach for training robotic manipulation policies directly in the physical world, a

EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins

SafetyDGX agent

arXiv:2607.08793v4 Announce Type: replace-cross Abstract: Sepsis is a leading cause of mortality, yet optimal treatment policies remain contested. Existing reinforcement learning (RL) approaches learn

Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning

SafetyDGX agent

arXiv:2608.09876v1 Announce Type: cross Abstract: Physically consistent motion planning remains a fundamental challenge in embodied AI, as generated trajectories must strictly conform to real-world ex

Ethical Framework for Responsible Foundational Models in Medical Imaging

SafetyDGX agent

arXiv:2406.11868v2 Announce Type: replace-cross Abstract: The emergence of foundational models represents a paradigm shift in medical imaging, offering extraordinary capabilities in disease detection,

Evaluation of Motivational Interviewing Counsellors with Task-Aware Multi-Stage LLM-Based Simulated Clients

SafetyDGX agent

arXiv:2608.07499v1 Announce Type: cross Abstract: The development and benchmarking of Large Language Model (LLM)-based Motivational Interviewing (MI) counsellors now often rely on LLM-based simulated

Evolving Safety Landscape of Multi-modal Large Language Models: A Survey of Emerging Threats and Safeguards

SafetyDGX agent

arXiv:2608.07535v1 Announce Type: cross Abstract: Multi-modal large language models (MLLMs) integrate heterogeneous modalities through modality alignment and fusion, enabling stronger understanding an

Explaining, Verifying, and Aligning Semantic Hierarchies in Vision-Language Model Embeddings

SafetyDGX agent

arXiv:2603.26798v2 Announce Type: replace-cross Abstract: Vision-language model (VLM) encoders such as CLIP enable strong retrieval and zero-shot classification in a shared image-text embedding space,

Explore, Map, Remember, Decide: Are Embodied VLMs Ready for Safety-Critical Scenarios?

SafetyDGX agent

arXiv:2608.08077v1 Announce Type: new Abstract: Theory of Space framework (ToS) assesses the spatial understanding of curiosity-driven Vision-Language Models (VLMs) under partial observability. As AI

← Previous
12345…210
Next →