AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
23 Jun 2026

FairBED: A Bayesian Experimental Design Approach to Gathering Fairer Data

SafetyDGX agent

arXiv:2606.23515v1 Announce Type: cross Abstract: Frameworks for ensuring fairness in machine learning typically focus on learning fair models from existing data. But this endeavor is often undermined

Fairness under Graph Uncertainty: Achieving Interventional Fairness with Partially Known Causal Graphs over Clusters of Variables

SafetyDGX agent

arXiv:2602.23611v2 Announce Type: replace-cross Abstract: Algorithmic decisions about individuals require predictions that are not only accurate but also fair with respect to sensitive attributes such

FairSAM: Fair Classification on Corrupted Image Data Through Sharpness-Aware Minimization

SafetyDGX agent

arXiv:2503.22934v2 Announce Type: replace Abstract: Image classification models trained on clean data often degrade sharply when exposed to corrupted test or deployment data, such as images with impul

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

FAIRVAR: Fair Federated Learning via Variance Regularization

SafetyDGX agent

arXiv:2508.12042v2 Announce Type: replace Abstract: Federated learning (FL) allows collaborative training of machine learning models across multiple parties without sharing raw data. However, heteroge

FAST: A Framework for Aligned Sampling and Training in Parallel Reinforcement Learning for Autonomous Driving

SafetyDGX agent

arXiv:2606.21587v1 Announce Type: new Abstract: Deep reinforcement learning is pivotal for closed-loop autonomous driving yet remains constrained by severe bottlenecks in sampling efficiency. Standard

Fast Nonparametric Conditional Independence Testing via Two-Stage Regression

SafetyDGX agent

arXiv:2606.18011v2 Announce Type: replace-cross Abstract: Constraint-based causal discovery relies on repeated conditional independence tests, but fast nonparametric tests often sacrifice calibration,

Federated Temporal Attention Intelligence for Cyber-Resilient IoMT: Lightweight Digital Twins and PPO-Driven Honeypot Deception

SafetyDGX agent

arXiv:2606.21422v1 Announce Type: new Abstract: The rapid proliferation of Internet of Medical Things (IoMT) devices introduces critical cybersecurity vulnerabilities in healthcare environments where

FILIC: Dual-Loop Force-Guided Imitation Learning with Impedance Torque Control for Contact-Rich Manipulation Tasks

SafetyDGX agent

arXiv:2509.17053v2 Announce Type: replace Abstract: Many contact-rich manipulation tasks require precise force regulation. However, most imitation learning (IL) policies remain position-centric and la

FlowDPG: Deterministic Policy Gradient on Flow Matching Policies for Real-World Manipulation

SafetyDGX agent

arXiv:2606.22303v1 Announce Type: new Abstract: Real-world reinforcement learning for robotic manipulation remains challenging, and this difficulty is amplified for flow matching policies: applying po

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

SafetyDGX agent

arXiv:2606.20867v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-sho

Formalizing Task-Space Complexity for Zero-Shot Generalization

SafetyDGX agent

arXiv:2606.20967v1 Announce Type: new Abstract: Policies must operate across diverse conditions, yet a single policy is often conservative while fully adaptive schemes can be complex. We study zero-sh

From Gradient Clipping to Structural Refinement: Improving DPSGD for Medical Image Segmentation

SafetyDGX agent

arXiv:2606.21763v1 Announce Type: new Abstract: Medical image segmentation is widely used for disease detection but relies on sensitive data, raising privacy concerns as trained models can leak inform

From Reconstruction to Decision: A Post-Encoder Plug-in Adapter for Curvilinear Segmentation

SafetyDGX agent

arXiv:2606.23486v1 Announce Type: new Abstract: Curvilinear object segmentation, including vessels and cracks, is challenging due to extreme spatial sparsity and topological fragility, where small loc

FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation

SafetyDGX agent

arXiv:2606.13102v2 Announce Type: replace Abstract: Despite the success of vision-based generalist robotic policies, existing tactile-based policies remain tied to fixed embodiments and sensor setups.

GAC: Stabilizing Asynchronous RL Training for LLMs via Gradient Alignment Control

SafetyDGX agent

arXiv:2603.01501v2 Announce Type: replace Abstract: Asynchronous execution is essential for scaling reinforcement learning (RL) to modern large model workloads, including large language models and AI

GARIP: A Running-Average Moving Reference for Last-Iterate Self-Play in Two-Player Zero-Sum Games

SafetyDGX agent

arXiv:2606.22688v1 Announce Type: cross Abstract: Self-play with naive gradient ascent cycles in two-player zero-sum games: the last iterate orbits the equilibrium. Modern methods restore last-iterate

Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems

SafetyDGX agent

arXiv:2505.00909v3 Announce Type: replace Abstract: In this paper, we propose a Gaussian Process (GP)-based policy iteration framework for addressing both forward and inverse problems in Hamilton--Jac

Geometric Entropy: When Trajectory Diversity Helps and Hurts in Imitation Learning

SafetyDGX agent

arXiv:2606.20871v1 Announce Type: new Abstract: We study how trajectory-shape diversity in demonstrations affects imitation learning (IL) performance across models, tasks, and data scales. We introduc

Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication

SafetyDGX agent

arXiv:2509.09597v3 Announce Type: replace-cross Abstract: Graph alignment, the problem of identifying corresponding nodes across multiple graphs, is fundamental to numerous applications. Most existing

Graph-of-Differences: Anatomy-Structured Difference Alignment for Medical Image Re-Identification

SafetyDGX agent

arXiv:2606.21368v1 Announce Type: new Abstract: Medical image re-identification (MedReID) enables longitudinal patient linkage but remains vulnerable to shortcut learning and often produces decisions

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2606.22995v1 Announce Type: new Abstract: Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy up

GTA-Net: Cooperative Game Theory for Vision-Language Alignment in Chest X-Ray Report Generation

SafetyDGX agent

arXiv:2606.21915v1 Announce Type: new Abstract: Automated chest X-ray report generation requires precise cross-modal grounding to ensure clinically reliable descriptions. However, existing vision-lang

Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation

SafetyDGX agent

arXiv:2602.00032v3 Announce Type: replace-cross Abstract: Synthetic faces from text-to-image (T2I) models pervade digital media, yet their demographic biases under emotionally conditioned prompts rema

HEAS: Hierarchical Evolutionary Agent-Based Simulation Framework for Multi-Objective Policy Search

SafetyDGX agent

arXiv:2508.15555v4 Announce Type: replace-cross Abstract: HEAS is a Python framework that connects agent-based simulation, evolutionary search, and scenario-based evaluation in a single reproducible p

Heterogeneous Policy Networks for Composite Robot Team Communication and Coordination

SafetyDGX agent

arXiv:2606.20962v1 Announce Type: new Abstract: High-performing human-human teams learn intelligent and efficient communication and coordination strategies to maximize their joint utility. These teams

Hierarchical Reinforcement Learning for Sparse-Reward Search in Commutative Algebra

SafetyDGX agent

arXiv:2606.22922v1 Announce Type: new Abstract: Applying machine learning techniques to solving long-standing mathematical conjectures can be particularly challenging due to their extreme reward spars

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning

SafetyDGX agent

arXiv:2606.22860v1 Announce Type: new Abstract: Recent advancements in generative imitation learning have significantly propelled the field of robotic manipulation. However, the majority of existing m

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

SafetyDGX agent

arXiv:2606.20189v2 Announce Type: replace Abstract: Leveraging Vision Foundation Models (VFMs) for camera-to-LiDAR knowledge distillation offers a promising solution to the scarcity of annotated data

Horizon Adaptive Offline Policy Learning via Value Stitching

SafetyDGX agent

arXiv:2606.21136v1 Announce Type: new Abstract: Learning accurate value functions plays a decisive role for reinforcement learning (RL) agents to solve long-horizon, complex tasks. Conventional tempor

Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation

SafetyDGX agent

arXiv:2509.17125v2 Announce Type: replace Abstract: Relational object rearrangement (ROR) tasks (e.g., insert flower to vase) require a robot to manipulate objects with precise semantic and geometric

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models

SafetyDGX agent

arXiv:2606.21672v1 Announce Type: cross Abstract: Imitation learning has emerged as a powerful paradigm for learning visuomotor policies, but its generalisation and stability are limited by the scale

Improved Algorithms for Nash Welfare in Linear Bandits

SafetyDGX agent

arXiv:2601.22969v2 Announce Type: replace Abstract: Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training

SafetyDGX agent

arXiv:2606.22158v1 Announce Type: new Abstract: Achieving human-like reasoning in Vision-Language Models (VLMs) remains a long-standing challenge. Recent approaches leverage Chain-of-Thought (CoT) rat

Improving Robotic Imitation Learning via Trajectory Standardization

SafetyDGX agent

arXiv:2606.22907v1 Announce Type: cross Abstract: Imitation learning for robotic manipulation relies on large sets of human demonstration trajectories, which are often noisy and temporally irregular d

Inductive Generalization for Robotic Manipulation

SafetyDGX agent

arXiv:2606.20999v1 Announce Type: cross Abstract: Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models l

Influencer Cartels

SafetyDGX agent

arXiv:2405.10231v3 Announce Type: replace-cross Abstract: Social media influencers account for a growing share of marketing worldwide. We demonstrate the existence of a novel form of market failure in

'Instead of allowing for greater focus, the latest AI tools are overwhelming workers, frazzling minds and shredding attention spans.' https:…

SafetyDGX agent

'Instead of allowing for greater focus, the latest AI tools are overwhelming workers, frazzling minds and shredding attention spans.' https://www.theatlantic.com/technology/2026/06/ai-agents-jobs-exha

Integrating Facial Generation into Full-Duplex Spoken Dialogue Systems

SafetyDGX agent

arXiv:2606.21970v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models, such as Moshi, enable natural, low-latency voice conversations. However, they remain limited to the audio modality

Interpretable Probabilistic Medical Image Segmentation via Gaussian Process with Explicit Modelling of Annotation Bias and Variability

SafetyDGX agent

arXiv:2606.23177v1 Announce Type: new Abstract: Deep learning-based medical image segmentation models are trained using annotations that exhibit systematic bias and variability across raters. While pr

Inverting the Bellman Equation: From Q-Values to World Models

SafetyDGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

IRumAI: Reinforcement Learning for Indian Rummy

SafetyDGX agent

arXiv:2606.21975v1 Announce Type: cross Abstract: Despite its massive player base and complex hidden-information dynamics, Indian Rummy has received no reinforcement learning attention. Existing agent

IViT: A Novel Interpretable Visual Transformer for Skin Disease Detection

SafetyDGX agent

arXiv:2606.22892v1 Announce Type: cross Abstract: The clinical diagnosis of skin diseases is susceptible to interference from inter-class similarity of skin lesions, and over-reliance on clinicians'ex

KITE: Decoupling Kinematics and Interaction for Zero-Shot Cross-Embodiment Manipulation

SafetyDGX agent

arXiv:2606.22113v1 Announce Type: new Abstract: Generalizing manipulation policies across robot embodiments remains difficult because standard policies entangle task reasoning with embodiment-specific

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation

SafetyDGX agent

arXiv:2606.23685v1 Announce Type: new Abstract: Human-hand demonstrations provide a direct and scalable source of physical interaction data for robot learning. While manual retargeting is indispensabl

Learned Controllers for Agile Quadrotors in Pursuit-Evasion Games

SafetyDGX agent

arXiv:2506.02849v3 Announce Type: replace-cross Abstract: In this letter we study 1v1 quadrotor pursuit-evasion, where a pursuer and an evader are trained via reinforcement learning (RL) by competing

Learning Control as Enabling Layer for Embodied Intelligence Research explored with Soft Robotic Swimming in diverse Flow Speeds

SafetyDGX agent

arXiv:2606.20660v1 Announce Type: new Abstract: Soft robots are valuable robophysical platforms for studying body-caudal undulatory locomotion, but their compliant bodies are difficult to control prec

Learning Process Rewards via Success Visitation Matching for Efficient RL

SafetyDGX agent

arXiv:2606.23640v1 Announce Type: new Abstract: In many modern applications of reinforcement learning (RL), the natural reward for a task of interest is inherently sparse: a reward of 0 is given every

Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem

SafetyDGX agent

arXiv:2606.21604v1 Announce Type: new Abstract: Neural combinatorial optimization (NCO) has shown that policies trained by reinforcement can construct strong solutions to NP-hard problems directly fro

Learning to See While Learning to Act: Diffusion Models for Active Perception in Robot Imitation

SafetyDGX agent

arXiv:2606.23625v1 Announce Type: new Abstract: Most imitation learning methods assume full observability in table-top settings. In practice, objects are often occluded, requiring robots to both searc

LOLLA: Deep Reinforcement Learning for Closed-Loop Link Adaptation Towards a GPU-Accelerated AI-RAN

SafetyDGX agent

arXiv:2606.23110v1 Announce Type: cross Abstract: Outer-loop link adaptation (OLLA) is widely deployed in 5G NR to track channel variations, yet its reliance on first-order, single-bit feedback degrad

Long-Distance Real-World Navigation of the Legged-Wheeled Robot Go2-W Using Deep Reinforcement Learning

SafetyDGX agent

arXiv:2606.21387v1 Announce Type: new Abstract: Legged-wheeled robots have long been studied for their potential to combine the efficient flat-ground mobility of wheels with the rough-terrain capabili

MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning

SafetyDGX agent

arXiv:2606.20641v1 Announce Type: cross Abstract: Multi-modal Large Language Models (MLLMs) have demonstrated remarkable capabilities in semantic understanding and common sense reasoning, making them

Maintain Plasticity in Long-timescale Continual Test-time Adaptation

SafetyDGX agent

arXiv:2412.20034v2 Announce Type: replace Abstract: Continual test-time domain adaptation (CTTA) aims to adjust pre-trained source models to perform well over time across non-stationary target environ

MAPS: Multi-Anchor Projection Similarity for Joint Vision-Language Geo-Localization

SafetyDGX agent

arXiv:2606.22543v1 Announce Type: new Abstract: Humans localize places by integrating perceptual cues from vision with semantic reasoning from language, forming a scene understanding that is both intu

Measuring Model-Induced Discrimination via Efficient Fairness Approximation

SafetyDGX agent

arXiv:2405.09251v2 Announce Type: replace Abstract: Providing various machine learning (ML) applications in the real world, concerns about discrimination hidden in ML models are growing, particularly

Memory Contagion: Cross-Temporal Propagation of Evaluator Bias via Agent Memory

SafetyDGX agent

arXiv:2606.23195v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly rely on memory systems to maintain long-term coherence. Recent work shows that agent memories degrade dur

MemoryVAM: Integrating Memory into Video Action Model for Robot Manipulation

SafetyDGX agent

arXiv:2606.20679v1 Announce Type: cross Abstract: Video-world-model policies learn action-relevant representations by predicting future observations. However, they condition on only a short observatio

Meta copied slot machines to addict kids to Instagram. Now Zuckerberg is turning his company into a prediction market. Meta’s business model…

SafetyDGX agent

Meta copied slot machines to addict kids to Instagram. Now Zuckerberg is turning his company into a prediction market. Meta’s business model is profiting from addiction—kids, gamblers, & more. Stop it

Meta-Reinforcement Learning via Evolution for Multi-Objective Combinatorial Supply Chain Optimisation

SafetyDGX agent

arXiv:2606.22146v1 Announce Type: new Abstract: Meta-reinforcement learning is a promising approach to multi-objective optimisation because it enables rapid policy adaptation across changing environme

Mind the Privileged-to-Camera Gap: Actor-Centric Sidecar Supervision for Camera-First Open-Loop Waypoint Prediction

SafetyDGX agent

arXiv:2606.20772v1 Announce Type: new Abstract: Camera-first autonomous-driving models predict future ego waypoints from images, ego-state features, and route commands, but waypoint supervision alone

← Previous
1…106107108109110…242
Next →