AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
23 Jun 2026

A Markov Chain Approach to Preference Alignment

SafetyDGX agent

arXiv:2606.22652v1 Announce Type: new Abstract: We propose Markov Chain from Human Feedback (MCHF), an elementary approach for aligning generative models from pairwise human preferences. Unlike Reinfo

A Neuromorphic Reinforcement Learning Framework for Efficient Pathfinding in Robotic Mobile Fulfillment Systems

SafetyDGX agent

arXiv:2606.20031v2 Announce Type: replace Abstract: Dynamic environmental changes, confined workspaces, and stringent real-time constraints make pathfinding in Robotic Mobile Fulfillment Systems (RMFS

A Stitch in Time Saves Nine: Preserving Policy Compatibility Under Perception Updates in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2606.21509v1 Announce Type: new Abstract: End-to-end autonomous driving systems tightly couple perception and decision-making through latent representations. Consequently, updates to perception


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A Taxonomy of Conceptual Alignment in Human-Robot Dialogue

SafetyDGX agent

arXiv:2606.22360v1 Announce Type: new Abstract: Successful conversations require speakers to align on the meaning of concepts, a challenging but crucial task for human-robot interaction. Understanding

A UAV-Based Multi-Modal Vision System for Automated Sideslope Deformation Monitoring and Hazard Detection

SafetyDGX agent

arXiv:2606.20681v1 Announce Type: new Abstract: Slope hazards constitute a major safety threat to expressway infrastructure, and their evolution is typically manifested as slow surface deformation. Co

A Watermark for Vision-Language-Action and World Action Models

SafetyDGX agent

arXiv:2606.23574v1 Announce Type: cross Abstract: Vision-language-action (VLA) models and world-action models (WAM) are the generative models now driving general-purpose robot control, turning raw cam

Action-BED: Task-Driven Bayesian Experimental Design with Singly Intractable Objectives

SafetyDGX agent

arXiv:2606.23662v1 Announce Type: cross Abstract: Bayesian experimental design (BED) has traditionally been based on maximising expected uncertainty reductions from prior to posterior. A major shortfa

Active Causal Experimentalist (ACE): Learning Intervention Strategies via Direct Preference Optimization

SafetyDGX agent

arXiv:2602.02451v2 Announce Type: replace Abstract: Discovering causal relationships requires controlled experiments, but experimentalists face a sequential decision problem: each intervention reveals

Adversarial observations in probabilistic State-Space Models for robust Reinforcement Learning

SafetyDGX agent

arXiv:2606.20880v1 Announce Type: cross Abstract: Decision-making under partial or adversarial observability requires accurate inference of the environment's latent state and its associated uncertaint

Again: data never stays secure and mass surveillance and data harvesting will always be a liability

SafetyDGX agent

Gary Marcus argues that data security is fundamentally fragile and that mass surveillance and data harvesting practices create persistent security vulnerabilities that cannot be eliminated. The post e

Agentic Time Machine as an Infrastructure for Future-Event Forecasting

SafetyDGX agent

arXiv:2606.21013v1 Announce Type: cross Abstract: Forecasting future events is a critical challenge for large language model (LLM) agents, spanning domains from elections and monetary policy to financ

AI Capex pushback from GS: 'If frontier intelligence can increasingly be developed in the East at a fraction of the cost incurred in the Wes…

SafetyDGX agent

AI Capex pushback from GS: 'If frontier intelligence can increasingly be developed in the East at a fraction of the cost incurred in the West … then the largest capital allocators are also the ones mo

AI data centres are hungry for land, water & power. I’m calling on every major AI company to publicly disclose the full environmental impact…

SafetyDGX agent

AI data centres are hungry for land, water & power. I’m calling on every major AI company to publicly disclose the full environmental impact of its systems – as a matter of transparency No more hidden

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training

SafetyDGX agent

arXiv:2602.12691v3 Announce Type: replace Abstract: We study how to improve large foundation vision-language-action (VLA) systems through human-in-the-loop reinforcement learning (RL) in real-world en

Although I supported Musk in his suit against OpenAI, and admire what he did for electric cars, I am afraid my considered overall view is no…

SafetyDGX agent

Although I supported Musk in his suit against OpenAI, and admire what he did for electric cars, I am afraid my considered overall view is not very different from this: Candidly I have no idea why anyo

An LLM-Explainable DRL Framework for Passenger-Directed Autonomous Driving

SafetyDGX agent

arXiv:2606.20640v1 Announce Type: cross Abstract: Autonomous vehicles offer the potential for safer and more efficient mobility, yet public trust remains limited due to the lack of transparency in the

Any-Body Guard: Universal Safeguarding for Manipulation Policies via Action Masking

SafetyDGX agent

arXiv:2606.22278v1 Announce Type: cross Abstract: Ensuring safety of learning-enabled robotic manipulation across diverse embodiments and tasks still requires significant manual engineering. Existing

APEX: Action Priors Enable Efficient Exploration for Robust Motion Tracking on Legged Robots

SafetyDGX agent

arXiv:2505.10022v4 Announce Type: replace Abstract: Learning natural, animal-like locomotion from demonstrations has become a core paradigm in legged robotics. While motion tracking can reproduce refe

ARGUSTRACK: A Multi-View Annotation System for Multi-Object Tracking

SafetyDGX agent

arXiv:2606.20687v1 Announce Type: new Abstract: Multi-Camera Multi-Target (MCMT) tracking has emerged as a critical capability for applications ranging from autonomous driving to animal behavior monit

ARP: Enhancing Quantized Skill Abstractions via Visual Alignment and Iterative Refinement for Robotic Manipulation

SafetyDGX agent

arXiv:2606.22480v1 Announce Type: new Abstract: Learning visuomotor policies for long-horizon manipulation remains a fundamental challenge. Recent skill-based imitation learning methods based on discr

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control

SafetyDGX agent

arXiv:2606.21525v1 Announce Type: new Abstract: Model-free reinforcement learning algorithms such as Proximal Policy Optimization (PPO) treat the environment as a black box, estimating policy gradient

BadDreamer: Transferable Backdoor Attacks against Video World Models for Autonomous Driving

SafetyDGX agent

arXiv:2606.21172v1 Announce Type: new Abstract: Video world models are increasingly used in autonomous driving to forecast future scene evolution and provide future-aware spatio-temporal representatio

Balancing Performance and Diversity in GRPO Autoregressive Text-to-Image Post-Training

SafetyDGX agent

arXiv:2606.21498v1 Announce Type: cross Abstract: Autoregressive text-to-image (T2I) generation has recently advanced rapidly, yet aligning generated images with human preferences remains challenging.

BARD-MARL: Byzantine-Agent Detection for Learned Communication in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.20701v1 Announce Type: cross Abstract: Learned communication improves coordination in cooperative multi-agent reinforcement learning, but it also creates a trust problem: a trained policy m

BayesFP: Posterior Estimation for Flow-Based Policies via Feynman-Kac Sampling

SafetyDGX agent

arXiv:2606.21014v1 Announce Type: new Abstract: Robots must generate trajectories that remain faithful to learned expert behavior while satisfying safety constraints and task-specific objectives speci

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models

SafetyDGX agent

arXiv:2606.21645v1 Announce Type: cross Abstract: Large language models (LLMs) can readily reproduce conventional expressions, yet their ability to model gradient frequency distributions remains under

B[FM]^2: Brain Foundation Model via Flow Matching with SplitUNet

SafetyDGX agent

arXiv:2606.20812v1 Announce Type: new Abstract: EEG foundation models can learn generalizable representations from large-scale EEG corpora to enable single-backbone transfer across diverse clinical an

BiliVLA: Scene-Aware Vision-Language-Action Model with Reinforcement Learning for Autonomous Biliary Endoscopic Navigation

SafetyDGX agent

arXiv:2606.23531v1 Announce Type: new Abstract: Endoscopic retrograde cholangiopancreatography (ERCP) demands precise endoscopic navigation and stable biliary cannulation within a narrow monocular fie

BLENDS: Bayesian Learning-Enhanced Deep Smoothing for GNSS-Denied Environments

SafetyDGX agent

arXiv:2606.22456v1 Announce Type: new Abstract: Maintaining accurate navigation during GNSS outages remains a significant challenge for autonomous systems relying on low-cost inertial sensors. While c

Boosting CVaR Policy Optimization with Quantile Gradients

SafetyDGX agent

arXiv:2601.22100v3 Announce Type: replace Abstract: Optimizing Conditional Value-at-risk (CVaR) using policy gradient (a.k.a CVaR-PG) faces significant challenges of sample inefficiency. This ineffici

Bridge the Gaps: Heterogeneous Attributed Graph Clustering via Quaternion Representation Learning

SafetyDGX agent

arXiv:2606.23199v1 Announce Type: new Abstract: Attributed graph clustering partitions nodes by jointly exploiting node attributes and graph topology. It remains challenging due to attribute heterogen

Bypassing Minimization Bias: A Shift-Invariant Variance Estimator for Off-Equilibrium Local Learning Coefficients

SafetyDGX agent

arXiv:2606.22389v1 Announce Type: new Abstract: Singular Learning Theory leverages the Local Learning Coefficient (LLC) to quantify the geometry of neural network loss landscapes. However, mean-energy

Can LLMs Control Readability? A Multi-Dimensional Evaluation Framework for CEFR-Controlled Arabic Generation

SafetyDGX agent

arXiv:2606.21981v1 Announce Type: cross Abstract: While Large Language Models (LLMs) can generate fluent Arabic text, their ability to reliably control readability levels remains unclear. We propose a

Causal Gaussian Processes for Robust Treatment Effect Evaluation with Unobserved Confounding

SafetyDGX agent

arXiv:2606.21809v1 Announce Type: new Abstract: The presence of confounding bias poses a key challenge in policy evaluation, as the target causal effects of actions are not identifiable (i.e., underde

Causally Fair Node Classification on Non-IID Graph Data

SafetyDGX agent

arXiv:2505.01652v2 Announce Type: replace Abstract: Fair machine learning seeks to identify and mitigate biases in predictions against unfavorable populations characterized by demographic attributes,

CFPO: Counterfactual Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2606.23206v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL)

Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL

SafetyDGX agent

arXiv:2602.03389v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning remains challenging for long-horizon tasks. While hierarchical approaches mitigate this issue by dec

CLAR: Learning 3D Representations for Robotic Manipulation by Fusing Masked Reconstruction with Multi-Level Contrastive Alignment

SafetyDGX agent

arXiv:2507.08262v2 Announce Type: replace-cross Abstract: The spatial information inherent in 3D point clouds is crucial for robotic manipulation. However, existing 3D pre-training methods face a fund

ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures

SafetyDGX agent

arXiv:2606.19380v2 Announce Type: replace-cross Abstract: Software engineering and deployment are increasingly delegated to AI coding agents. The scale of their adoption is surfacing rare, but highly

Closed-Loop Verbal Reinforcement Learning for Task-Level Robotic Planning

SafetyDGX agent

arXiv:2603.22169v2 Announce Type: replace Abstract: We propose a new Verbal Reinforcement Learning (VRL) framework for interpretable task-level planning in mobile robotic systems operating under execu

Cohort-Anchored Foundation Models for Electronic Health Records: From Risk Scores to Auditable Peer Cohorts

SafetyDGX agent

arXiv:2606.21885v1 Announce Type: new Abstract: Foundation models have achieved remarkable performance across medical question answering, imaging, and electronic health record (EHR) tasks, yet reliabl

Concept Alignment Contrast and Long-Short Prompt Memory for Test-Time Adaptation of SAM3 in Medical Image Segmentation

SafetyDGX agent

arXiv:2606.22963v1 Announce Type: new Abstract: Concept segmentation models like Segment Anything Model 3 (SAM3) show strong generalization on natural images, yet their performance degrades in medical

Conditional Flow Matching for Visually-Guided Acoustic Highlighting

SafetyDGX agent

arXiv:2602.03762v3 Announce Type: replace-cross Abstract: Visually-guided acoustic highlighting seeks to rebalance audio in alignment with the accompanying video, creating a coherent audio-visual expe

Confidence-Uncertainty Boundary Calibration for Bayesian Deep Learning in Medical Image Analysis

SafetyDGX agent

arXiv:2602.11973v2 Announce Type: replace Abstract: In critical decision support systems based on medical imaging, the reliability of AI-assisted decision-making is as relevant as predictive accuracy.

Conflict-Aware Switching for CBF-CLF-Based Multi-Goal Navigation

SafetyDGX agent

arXiv:2606.21577v1 Announce Type: new Abstract: Quadratic programs (QPs) using Control Barrier Functions (CBFs) and Control Lyapunov Functions (CLFs) are widely used for safe control in reach-and-avoi

Continuous Behavioral Authentication via Multi-Expert BERT Log Analysis for Secure Data Sharing

SafetyDGX agent

arXiv:2606.21900v1 Announce Type: cross Abstract: Continuous authentication for mobile and zero-trust systems requires nonintrusive evidence confirming the enrolled user-device context remains valid a

Continuous-Time Probabilistic Correctors for Uncertainty-Aware Physics-Based Spacecraft Trajectory Forecasting

SafetyDGX agent

arXiv:2606.21021v1 Announce Type: new Abstract: Long-horizon spacecraft trajectory forecasting suffers from error accumulation due to the absence of corrective observations in the forecast regime, mak

CoorDex: Coordinating Body and Hand Priors for Continuous Dexterous Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.23680v1 Announce Type: cross Abstract: Humanoid loco-manipulation is often simplified into a stop-and-go process: walking to an object, stopping to manipulate it, and then resuming locomoti

Counsel: A Meta-Evaluation Dataset for Agentic Tasks

SafetyDGX agent

arXiv:2606.21627v1 Announce Type: cross Abstract: As agentic systems tackle increasingly complex multi-step tasks, evaluating their trajectories presents a major bottleneck - human annotation of a sin

Counterfactual learning of new adaptive instructional policies using logged data

SafetyDGX agent

arXiv:2606.23015v1 Announce Type: new Abstract: Optimizing instructional policies in Intelligent Tutoring Systems (ITS) typically requires costly online experimentation or student simulators that may

Cross-Modal Corroboration for Annotation-Free Wildlife Monitoring

SafetyDGX agent

arXiv:2606.21613v1 Announce Type: new Abstract: Scaling wildlife monitoring for real-world conservation deployments requires automated analysis of smart sensors that operate under severe annotation sc

Customizing Video Portraits via Identity-ActionDecoupling

SafetyDGX agent

arXiv:2606.22347v1 Announce Type: new Abstract: Identity-Preserving Text-to-Video Generation (IPT2V) seeks to synthesize a temporally coherent video from a reference image and a textual description, w

D2HDMap: Non-visible Driveline Map Prior for Online Vectorized HD Map Prediction

SafetyDGX agent

arXiv:2606.20725v1 Announce Type: new Abstract: Accurate, up-to-date representations of road structures are critical for the safe operation of autonomous vehicles. Existing systems rely either on cost

Darwin Mobile Agent: A Roadmap for Self-Evolution

SafetyDGX agent

arXiv:2606.20622v1 Announce Type: cross Abstract: The goal of artificial intelligence is to create agents capable of general, adaptive behaviour in open-ended environments. Guided by the 'Bitter Lesso

DASIP: Dynamic Test-Time Compute Scaling for Robot Control with Stochastic Interpolant Policies

SafetyDGX agent

arXiv:2511.20906v2 Announce Type: replace Abstract: Diffusion- and flow-based policies deliver state-of-the-art performance on long-horizon robotic manipulation and imitation learning tasks. However,

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

SafetyDGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

DataMIL: Selecting Data for Robot Imitation Learning with Datamodels

SafetyDGX agent

arXiv:2505.09603v2 Announce Type: replace-cross Abstract: Recently, the robotics community has amassed ever larger and more diverse datasets to train generalist policies. However, while these policies

DBT-Bleed: Dual-Branch Temporal Modeling with Key-Frame Selection for Surgical Bleeding Detection

SafetyDGX agent

arXiv:2606.22829v1 Announce Type: new Abstract: Intraoperative Adverse Events (IAEs) detection is critical for improving surgical safety, with bleeding being among the most frequent events across many

Deep Learning for Individual Heterogeneity

SafetyDGX agent

arXiv:2010.14694v4 Announce Type: replace-cross Abstract: This paper integrates deep neural networks (DNNs) into structural models to increase flexibility and capture rich heterogeneity while preservi

Deep RL for Fast Long-Horizon Operations Scheduling on NASA's Carruthers Geocorona Observatory Mission

SafetyDGX agent

arXiv:2606.22159v1 Announce Type: cross Abstract: Spacecraft operations scheduling is a highly constrained, long-horizon combinatorial optimization problem that traditionally relies on heuristics, con

← Previous
1…6768697071…214
Next →