AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
23 Jun 2026

DeformX: A Versatile Co-Simulation Framework for Deformable Linear Objects

SafetyDGX agent

arXiv:2606.22116v1 Announce Type: new Abstract: Deformable linear objects (DLOs) such as wires, cables, and ropes are common in robotic manipulation tasks, yet simulating them with both visual realism

Delta-Diffusion: Modeling Longitudinal Brain Amyloid-PET Trajectories via Conditional Poisson Diffusion Bridge

SafetyDGX agent

arXiv:2606.22216v1 Announce Type: cross Abstract: While longitudinal brain PET imaging is the gold standard for quantifying the spatiotemporal accumulation of Beta-amyloid, its widespread clinical uti

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views

SafetyDGX agent

arXiv:2606.23557v1 Announce Type: new Abstract: Multi-view 3D Visual Question Answering (MV3D-VQA) requires integrating partial observations into a coherent 3D scene representation and selecting infor


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Distribution-Aware Diffusion-LLM for Robust Ultra-Long-Term Time Series Forecasting

SafetyDGX agent

arXiv:2606.23391v1 Announce Type: new Abstract: Time series forecasting is a fundamental machine learning task. Recent work has explored Large Language Models (LLMs) for this purpose due to their stro

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling

SafetyDGX agent

arXiv:2606.23626v1 Announce Type: new Abstract: Can representations learned for image generation also support the evaluation of generated images? We study text-to-image reward prediction as a downstre

Do Activation Monitors Survive Model Updates? Benchmarking, Predicting, and Repairing Activation-Monitor Staleness

SafetyDGX agent

arXiv:2606.15980v2 Announce Type: replace Abstract: Activation monitors -- lightweight probes trained on a language model's internal representations -- are an increasingly common layer in deployment s

Don't Tell the Answer, Truly Guide the Reasoning During RL Rollouts

SafetyDGX agent

arXiv:2510.09388v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has become a key driver for enhancing the long chain-of-thought (CoT) reasoning capabilities of Large Language Models (L

DT-GOL: Dual-Track Geometric Online Learning in Nonstationary Environment with Label Delay

SafetyDGX agent

arXiv:2606.22950v1 Announce Type: new Abstract: Online learning is crucial for handling complex data streams in big data applications. Recent research has begun to focus on dynamic scenarios, i.e., no

Dual-Attention Convolution Experts for Sparse Tensor Completion

SafetyDGX agent

arXiv:2606.21427v1 Announce Type: new Abstract: Tensor factorization (TF) has been widely adopted for high-dimensional sparse data completion tasks. Despite significant progress, neural TF methods oft

dVLA-RL: Reinforcement Learning over Denoising Trajectories for Discrete Diffusion Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.23623v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have established a powerful paradigm for generalist robotic manipulation by grounding control into the semantic reas

EasyUUV: An LLM-Enhanced Universal and Lightweight Sim-to-Real Reinforcement Learning Framework for UUV Attitude Control

SafetyDGX agent

arXiv:2510.22126v3 Announce Type: replace Abstract: Despite recent advances in Unmanned Underwater Vehicle (UUV) attitude control, existing methods still struggle with generalizability, robustness to

Efficient Reinforcement Finetuning via Adaptive Curriculum Learning

SafetyDGX agent

arXiv:2504.05520v4 Announce Type: replace Abstract: Reinforcement finetuning (RFT) has shown great potential for enhancing the mathematical reasoning capabilities of large language models (LLMs), but

Efficient Training of Boltzmann Generators Using Off-Policy Log-Dispersion Regularization

SafetyDGX agent

arXiv:2602.03729v2 Announce Type: replace Abstract: Sampling from unnormalized probability densities is a central challenge in computational science. Boltzmann generators are generative models that en

EmbodiedUS-FS: Fast Slow Intelligence for Ultrasound Robotics

SafetyDGX agent

arXiv:2606.22319v1 Announce Type: cross Abstract: Robotic ultrasound scanning in real clinical environments requires both high-level clinical workflow reasoning and low-level closed-loop execution. Ph

Encoder-Decoder Manifold Alignment for Idempotent Generation

SafetyDGX agent

arXiv:2606.22304v1 Announce Type: new Abstract: Recently, several learning paradigms have been introduced to enforce idempotency in generative models. The goal is to ensure that repeated application o

Enhancing IMU-Based Online Handwriting Recognition via Contrastive Learning with Zero Inference Overhead

SafetyDGX agent

arXiv:2602.07049v2 Announce Type: replace Abstract: Online handwriting recognition using inertial measurement units opens up handwriting on paper as input for digital devices. Doing it on edge hardwar

Enhancing LLMs for Graph Tasks via Graph-aware LoRA Generation

SafetyDGX agent

arXiv:2606.22429v1 Announce Type: new Abstract: Graph neural networks (GNNs) tightly couple their input-output parameters to dataset-specific feature spaces and target sets, exhibiting limited transfe

Enhancing Road Safety: An IoT-Based Accident Detection and Prevention Mechanism

SafetyDGX agent

arXiv:2606.22381v1 Announce Type: cross Abstract: Road traffic accidents remain a critical global crisis, consistently serving as a primary driver of preventable mortality and severe injury. These inc

EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors

SafetyDGX agent

arXiv:2602.21218v2 Announce Type: replace-cross Abstract: High-quality data is essential for modern machine learning, yet many valuable corpora are sensitive and cannot be freely shared. Synthetic dat

EvoRubrics: Dynamic Rubrics as Rewards via Adversarial Co-Evolution for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.23038v1 Announce Type: new Abstract: Rubric-based rewards offer interpretable and fine-grained optimization signals for reinforcement learning in open-ended tasks where verifiable answers a

Expected Free Energy-based Planning as Variational Inference

SafetyDGX agent

arXiv:2606.20658v1 Announce Type: cross Abstract: Planning under uncertainty requires agents to balance goal achievement with information gathering. Active inference addresses this through the Expecte

Expert Consensus on Criteria for the Automated Assessment of Laparoscopic Camera Navigation

SafetyDGX agent

arXiv:2606.23131v1 Announce Type: new Abstract: Background: Laparoscopic camera navigation (LCN) is a critical skill, yet its current assessment typically relies on manual rating systems which are tim

Extraction and Analysis of Multimodal Concepts in Vision Language Models through Sparse Autoencoders

SafetyDGX agent

arXiv:2606.21197v1 Announce Type: new Abstract: Vision Language Models (VLMs) have demonstrated impressive performance in tasks requiring joint understanding of images and text, such as image captioni

Fair Transit Stop Placement: A Clustering Perspective and Beyond

SafetyDGX agent

arXiv:2602.06776v2 Announce Type: replace-cross Abstract: We study the transit stop placement (TrSP) problem in general metric spaces, where agents travel between source-destination pairs and may eith

FairBED: A Bayesian Experimental Design Approach to Gathering Fairer Data

SafetyDGX agent

arXiv:2606.23515v1 Announce Type: cross Abstract: Frameworks for ensuring fairness in machine learning typically focus on learning fair models from existing data. But this endeavor is often undermined

Fairness under Graph Uncertainty: Achieving Interventional Fairness with Partially Known Causal Graphs over Clusters of Variables

SafetyDGX agent

arXiv:2602.23611v2 Announce Type: replace-cross Abstract: Algorithmic decisions about individuals require predictions that are not only accurate but also fair with respect to sensitive attributes such

FairSAM: Fair Classification on Corrupted Image Data Through Sharpness-Aware Minimization

SafetyDGX agent

arXiv:2503.22934v2 Announce Type: replace Abstract: Image classification models trained on clean data often degrade sharply when exposed to corrupted test or deployment data, such as images with impul

FAIRVAR: Fair Federated Learning via Variance Regularization

SafetyDGX agent

arXiv:2508.12042v2 Announce Type: replace Abstract: Federated learning (FL) allows collaborative training of machine learning models across multiple parties without sharing raw data. However, heteroge

FAST: A Framework for Aligned Sampling and Training in Parallel Reinforcement Learning for Autonomous Driving

SafetyDGX agent

arXiv:2606.21587v1 Announce Type: new Abstract: Deep reinforcement learning is pivotal for closed-loop autonomous driving yet remains constrained by severe bottlenecks in sampling efficiency. Standard

Fast Nonparametric Conditional Independence Testing via Two-Stage Regression

SafetyDGX agent

arXiv:2606.18011v2 Announce Type: replace-cross Abstract: Constraint-based causal discovery relies on repeated conditional independence tests, but fast nonparametric tests often sacrifice calibration,

Federated Temporal Attention Intelligence for Cyber-Resilient IoMT: Lightweight Digital Twins and PPO-Driven Honeypot Deception

SafetyDGX agent

arXiv:2606.21422v1 Announce Type: new Abstract: The rapid proliferation of Internet of Medical Things (IoMT) devices introduces critical cybersecurity vulnerabilities in healthcare environments where

FILIC: Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks

SafetyDGX agent

arXiv:2509.17053v2 Announce Type: replace Abstract: Many contact-rich manipulation tasks require precise force regulation. However, most imitation learning (IL) policies remain position-centric and la

FlowDPG: Deterministic Policy Gradient on Flow Matching Policies for Real-World Manipulation

SafetyDGX agent

arXiv:2606.22303v1 Announce Type: new Abstract: Real-world reinforcement learning for robotic manipulation remains challenging, and this difficulty is amplified for flow matching policies: applying po

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

SafetyDGX agent

arXiv:2606.20867v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-sho

Formalizing Task-Space Complexity for Zero-Shot Generalization

SafetyDGX agent

arXiv:2606.20967v1 Announce Type: new Abstract: Policies must operate across diverse conditions, yet a single policy is often conservative while fully adaptive schemes can be complex. We study zero-sh

From Driving Videos to Simulatable Scenarios

SafetyDGX agent

arXiv:2606.21993v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) face driving scenarios ranging from routine traffic to rare events. To assess safety it is crucial to reproduce these scenar

From Gradient Clipping to Structural Refinement: Improving DPSGD for Medical Image Segmentation

SafetyDGX agent

arXiv:2606.21763v1 Announce Type: new Abstract: Medical image segmentation is widely used for disease detection but relies on sensitive data, raising privacy concerns as trained models can leak inform

From Reconstruction to Decision: A Post-Encoder Plug-in Adapter for Curvilinear Segmentation

SafetyDGX agent

arXiv:2606.23486v1 Announce Type: new Abstract: Curvilinear object segmentation, including vessels and cracks, is challenging due to extreme spatial sparsity and topological fragility, where small loc

FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation

SafetyDGX agent

arXiv:2606.13102v2 Announce Type: replace Abstract: Despite the success of vision-based generalist robotic policies, existing tactile-based policies remain tied to fixed embodiments and sensor setups.

GAC: Stabilizing Asynchronous RL Training for LLMs via Gradient Alignment Control

SafetyDGX agent

arXiv:2603.01501v2 Announce Type: replace Abstract: Asynchronous execution is essential for scaling reinforcement learning (RL) to modern large model workloads, including large language models and AI

GARIP: A Running-Average Moving Reference for Last-Iterate Self-Play in Two-Player Zero-Sum Games

SafetyDGX agent

arXiv:2606.22688v1 Announce Type: cross Abstract: Self-play with naive gradient ascent cycles in two-player zero-sum games: the last iterate orbits the equilibrium. Modern methods restore last-iterate

Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems

SafetyDGX agent

arXiv:2505.00909v3 Announce Type: replace Abstract: In this paper, we propose a Gaussian Process (GP)-based policy iteration framework for addressing both forward and inverse problems in Hamilton--Jac

Geometric Entropy: When Trajectory Diversity Helps and Hurts in Imitation Learning

SafetyDGX agent

arXiv:2606.20871v1 Announce Type: new Abstract: We study how trajectory-shape diversity in demonstrations affects imitation learning (IL) performance across models, tasks, and data scales. We introduc

Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication

SafetyDGX agent

arXiv:2509.09597v3 Announce Type: replace-cross Abstract: Graph alignment, the problem of identifying corresponding nodes across multiple graphs, is fundamental to numerous applications. Most existing

Graph-of-Differences: Anatomy-Structured Difference Alignment for Medical Image Re-Identification

SafetyDGX agent

arXiv:2606.21368v1 Announce Type: new Abstract: Medical image re-identification (MedReID) enables longitudinal patient linkage but remains vulnerable to shortcut learning and often produces decisions

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.22995v1 Announce Type: new Abstract: Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy up

GTA-Net: Cooperative Game Theory for Vision-Language Alignment in Chest X-Ray Report Generation

SafetyDGX agent

arXiv:2606.21915v1 Announce Type: new Abstract: Automated chest X-ray report generation requires precise cross-modal grounding to ensure clinically reliable descriptions. However, existing vision-lang

Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation

SafetyDGX agent

arXiv:2602.00032v3 Announce Type: replace-cross Abstract: Synthetic faces from text-to-image (T2I) models pervade digital media, yet their demographic biases under emotionally conditioned prompts rema

HEAS: Hierarchical Evolutionary Agent-Based Simulation Framework for Multi-Objective Policy Search

SafetyDGX agent

arXiv:2508.15555v4 Announce Type: replace-cross Abstract: HEAS is a Python framework that connects agent-based simulation, evolutionary search, and scenario-based evaluation in a single reproducible p

Helping build shared standards for advanced AI

SafetyDGX agent

OpenAI discusses its efforts to contribute to the development of shared industry standards and best practices for advanced artificial intelligence systems. The article likely covers OpenAI's involveme

Heterogeneous Policy Networks for Composite Robot Team Communication and Coordination

SafetyDGX agent

arXiv:2606.20962v1 Announce Type: new Abstract: High-performing human-human teams learn intelligent and efficient communication and coordination strategies to maximize their joint utility. These teams

Hierarchical Reinforcement Learning for Sparse-Reward Search in Commutative Algebra

SafetyDGX agent

arXiv:2606.22922v1 Announce Type: new Abstract: Applying machine learning techniques to solving long-standing mathematical conjectures can be particularly challenging due to their extreme reward spars

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning

SafetyDGX agent

arXiv:2606.22860v1 Announce Type: new Abstract: Recent advancements in generative imitation learning have significantly propelled the field of robotic manipulation. However, the majority of existing m

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

SafetyDGX agent

arXiv:2606.20189v2 Announce Type: replace Abstract: Leveraging Vision Foundation Models (VFMs) for camera-to-LiDAR knowledge distillation offers a promising solution to the scarcity of annotated data

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

SafetyDGX agent

arXiv:2606.23565v1 Announce Type: cross Abstract: LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise a

Horizon Adaptive Offline Policy Learning via Value Stitching

SafetyDGX agent

arXiv:2606.21136v1 Announce Type: new Abstract: Learning accurate value functions plays a decisive role for reinforcement learning (RL) agents to solve long-horizon, complex tasks. Conventional tempor

HumanHalo -- Safe and Efficient 3D Navigation Among Humans via Minimally Conservative MPC

SafetyDGX agent

arXiv:2510.17525v3 Announce Type: replace Abstract: Safe and efficient robotic navigation among humans is essential for integrating robots into everyday environments. Most existing approaches focus on

Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation

SafetyDGX agent

arXiv:2509.17125v2 Announce Type: replace Abstract: Relational object rearrangement (ROR) tasks (e.g., insert flower to vase) require a robot to manipulate objects with precise semantic and geometric

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models

SafetyDGX agent

arXiv:2606.21672v1 Announce Type: cross Abstract: Imitation learning has emerged as a powerful paradigm for learning visuomotor policies, but its generalisation and stability are limited by the scale

Improved Algorithms for Nash Welfare in Linear Bandits

SafetyDGX agent

arXiv:2601.22969v2 Announce Type: replace Abstract: Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social

← Previous
1…6869707172…214
Next →