AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
Safety

FOCA: Future-Oriented Conditioning for Data-Efficient Vision-Language-Action Adaptation

DGX agent

arXiv:2606.20867v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models enable general-purpose robotic control via large-scale multimodal pretraining, yet their effectiveness under few-sho

safetyarxiv-cs-cv
23 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Formalizing Task-Space Complexity for Zero-Shot Generalization

DGX agent

arXiv:2606.20967v1 Announce Type: new Abstract: Policies must operate across diverse conditions, yet a single policy is often conservative while fully adaptive schemes can be complex. We study zero-sh

safetyarxiv-cs-lg
23 Jun 2026
Safety

From Driving Videos to Simulatable Scenarios

DGX agent

arXiv:2606.21993v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) face driving scenarios ranging from routine traffic to rare events. To assess safety it is crucial to reproduce these scenar

safetyarxiv-cs-cv
23 Jun 2026
Safety

From Gradient Clipping to Structural Refinement: Improving DPSGD for Medical Image Segmentation

DGX agent

arXiv:2606.21763v1 Announce Type: new Abstract: Medical image segmentation is widely used for disease detection but relies on sensitive data, raising privacy concerns as trained models can leak inform

safetyarxiv-cs-cv
23 Jun 2026
Safety

From Reconstruction to Decision: A Post-Encoder Plug-in Adapter for Curvilinear Segmentation

DGX agent

arXiv:2606.23486v1 Announce Type: new Abstract: Curvilinear object segmentation, including vessels and cracks, is challenging due to extreme spatial sparsity and topological fragility, where small loc

safetyarxiv-cs-cv
23 Jun 2026
Safety

FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation

DGX agent

arXiv:2606.13102v2 Announce Type: replace Abstract: Despite the success of vision-based generalist robotic policies, existing tactile-based policies remain tied to fixed embodiments and sensor setups.

safetyarxiv-cs-ro
23 Jun 2026
Safety

GAC: Stabilizing Asynchronous RL Training for LLMs via Gradient Alignment Control

DGX agent

arXiv:2603.01501v2 Announce Type: replace Abstract: Asynchronous execution is essential for scaling reinforcement learning (RL) to modern large model workloads, including large language models and AI

safetyarxiv-cs-lg
23 Jun 2026
Safety

GARIP: A Running-Average Moving Reference for Last-Iterate Self-Play in Two-Player Zero-Sum Games

DGX agent

arXiv:2606.22688v1 Announce Type: cross Abstract: Self-play with naive gradient ascent cycles in two-player zero-sum games: the last iterate orbits the equilibrium. Modern methods restore last-iterate

safetyarxiv-cs-lg
23 Jun 2026
Safety

Gaussian process policy iteration with additive Schwarz acceleration for forward and inverse HJB and mean field game problems

DGX agent

arXiv:2505.00909v3 Announce Type: replace Abstract: In this paper, we propose a Gaussian Process (GP)-based policy iteration framework for addressing both forward and inverse problems in Hamilton--Jac

safetyarxiv-cs-lg
23 Jun 2026
Safety

Geometric Entropy: When Trajectory Diversity Helps and Hurts in Imitation Learning

DGX agent

arXiv:2606.20871v1 Announce Type: new Abstract: We study how trajectory-shape diversity in demonstrations affects imitation learning (IL) performance across models, tasks, and data scales. We introduc

safetyarxiv-cs-ro
23 Jun 2026
Safety

Graph Alignment via Dual-Pass Spectral Encoding and Latent Space Communication

DGX agent

arXiv:2509.09597v3 Announce Type: replace-cross Abstract: Graph alignment, the problem of identifying corresponding nodes across multiple graphs, is fundamental to numerous applications. Most existing

safetyarxiv-cs-cv
23 Jun 2026
Safety

Graph-of-Differences: Anatomy-Structured Difference Alignment for Medical Image Re-Identification

DGX agent

arXiv:2606.21368v1 Announce Type: new Abstract: Medical image re-identification (MedReID) enables longitudinal patient linkage but remains vulnerable to shortcut learning and often produces decisions

safetyarxiv-cs-cv
23 Jun 2026
Safety

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

DGX agent

arXiv:2606.22995v1 Announce Type: new Abstract: Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy up

safetyarxiv-cs-lg
23 Jun 2026
Safety

GTA-Net: Cooperative Game Theory for Vision-Language Alignment in Chest X-Ray Report Generation

DGX agent

arXiv:2606.21915v1 Announce Type: new Abstract: Automated chest X-ray report generation requires precise cross-modal grounding to ensure clinically reliable descriptions. However, existing vision-lang

safetyarxiv-cs-cv
23 Jun 2026
Safety

Happy Young Women, Grumpy Old Men? Emotion-Driven Demographic Biases in Synthetic Face Generation

DGX agent

arXiv:2602.00032v3 Announce Type: replace-cross Abstract: Synthetic faces from text-to-image (T2I) models pervade digital media, yet their demographic biases under emotionally conditioned prompts rema

safetyarxiv-cs-cv
23 Jun 2026
Safety

HEAS: Hierarchical Evolutionary Agent-Based Simulation Framework for Multi-Objective Policy Search

DGX agent

arXiv:2508.15555v4 Announce Type: replace-cross Abstract: HEAS is a Python framework that connects agent-based simulation, evolutionary search, and scenario-based evaluation in a single reproducible p

safetyarxiv-cs-lg
23 Jun 2026
Safety

Helping build shared standards for advanced AI

DGX agent

OpenAI discusses its efforts to contribute to the development of shared industry standards and best practices for advanced artificial intelligence systems. The article likely covers OpenAI's involveme

safetyopenai
23 Jun 2026
Safety

Heterogeneous Policy Networks for Composite Robot Team Communication and Coordination

DGX agent

arXiv:2606.20962v1 Announce Type: new Abstract: High-performing human-human teams learn intelligent and efficient communication and coordination strategies to maximize their joint utility. These teams

safetyarxiv-cs-ro
23 Jun 2026
Safety

Hierarchical Reinforcement Learning for Sparse-Reward Search in Commutative Algebra

DGX agent

arXiv:2606.22922v1 Announce Type: new Abstract: Applying machine learning techniques to solving long-standing mathematical conjectures can be particularly challenging due to their extreme reward spars

safetyarxiv-cs-lg
23 Jun 2026
Safety

HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning

DGX agent

arXiv:2606.22860v1 Announce Type: new Abstract: Recent advancements in generative imitation learning have significantly propelled the field of robotic manipulation. However, the majority of existing m

safetyarxiv-cs-ro
23 Jun 2026
Safety

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

DGX agent

arXiv:2606.20189v2 Announce Type: replace Abstract: Leveraging Vision Foundation Models (VFMs) for camera-to-LiDAR knowledge distillation offers a promising solution to the scarcity of annotated data

safetyarxiv-cs-cv
23 Jun 2026
Safety

HoloAgent-0: A Unified Embodied Agent Framework with 3D Spatial Memory

DGX agent

arXiv:2606.23565v1 Announce Type: cross Abstract: LLM agents follow a practical execution loop in digital environments: they reason over structured states, invoke tools, inspect feedback, and revise a

safetyarxiv-cs-cv
23 Jun 2026
Safety

Horizon Adaptive Offline Policy Learning via Value Stitching

DGX agent

arXiv:2606.21136v1 Announce Type: new Abstract: Learning accurate value functions plays a decisive role for reinforcement learning (RL) agents to solve long-horizon, complex tasks. Conventional tempor

safetyarxiv-cs-lg
23 Jun 2026
Safety

HumanHalo -- Safe and Efficient 3D Navigation Among Humans via Minimally Conservative MPC

DGX agent

arXiv:2510.17525v3 Announce Type: replace Abstract: Safe and efficient robotic navigation among humans is essential for integrating robots into everyday environments. Most existing approaches focus on

safetyarxiv-cs-ro
23 Jun 2026
Safety

Imagine2Act: Leveraging Object-Action Motion Consistency from Imagined Goals for Robotic Manipulation

DGX agent

arXiv:2509.17125v2 Announce Type: replace Abstract: Relational object rearrangement (ROR) tasks (e.g., insert flower to vase) require a robot to manipulate objects with precise semantic and geometric

safetyarxiv-cs-ro
23 Jun 2026
Safety

Imitation from Heterogeneous Demonstrations using Grounded Latent-Action World Models

DGX agent

arXiv:2606.21672v1 Announce Type: cross Abstract: Imitation learning has emerged as a powerful paradigm for learning visuomotor policies, but its generalisation and stability are limited by the scale

safetyarxiv-cs-lg
23 Jun 2026
Safety

Improved Algorithms for Nash Welfare in Linear Bandits

DGX agent

arXiv:2601.22969v2 Announce Type: replace Abstract: Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social

safetyarxiv-cs-lg
23 Jun 2026
Safety

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training

DGX agent

arXiv:2606.22158v1 Announce Type: new Abstract: Achieving human-like reasoning in Vision-Language Models (VLMs) remains a long-standing challenge. Recent approaches leverage Chain-of-Thought (CoT) rat

safetyarxiv-cs-cv
23 Jun 2026
Safety

Improving Robotic Imitation Learning via Trajectory Standardization

DGX agent

arXiv:2606.22907v1 Announce Type: cross Abstract: Imitation learning for robotic manipulation relies on large sets of human demonstration trajectories, which are often noisy and temporally irregular d

safetyarxiv-cs-cv
23 Jun 2026
Safety

Inductive Generalization for Robotic Manipulation

DGX agent

arXiv:2606.20999v1 Announce Type: cross Abstract: Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models l

safetyarxiv-cs-lg
23 Jun 2026
Safety

Influencer Cartels

DGX agent

arXiv:2405.10231v3 Announce Type: replace-cross Abstract: Social media influencers account for a growing share of marketing worldwide. We demonstrate the existence of a novel form of market failure in

safetyarxiv-cs-lg
23 Jun 2026
Safety

'Instead of allowing for greater focus, the latest AI tools are overwhelming workers, frazzling minds and shredding attention spans.' https:…

DGX agent

'Instead of allowing for greater focus, the latest AI tools are overwhelming workers, frazzling minds and shredding attention spans.' https://www.theatlantic.com/technology/2026/06/ai-agents-jobs-exha

safetygary-marcus--x
23 Jun 2026
Safety

Integrating Facial Generation into Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2606.21970v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models, such as Moshi, enable natural, low-latency voice conversations. However, they remain limited to the audio modality

safetyarxiv-cs-cv
23 Jun 2026
Safety

Intent-Handover: Grounding Language in Human-Usage Regions for Trustworthy Robot-to-Human Handovers

DGX agent

arXiv:2503.03579v2 Announce Type: replace-cross Abstract: Spoken instructions in robot-to-human handovers may specify either an object ('the cup') or an intended use ('pour water'); in both cases, suc

safetyarxiv-cs-lg
23 Jun 2026
Safety

Interpretable Probabilistic Medical Image Segmentation via Gaussian Process with Explicit Modelling of Annotation Bias and Variability

DGX agent

arXiv:2606.23177v1 Announce Type: new Abstract: Deep learning-based medical image segmentation models are trained using annotations that exhibit systematic bias and variability across raters. While pr

safetyarxiv-cs-cv
23 Jun 2026
Safety

Inverting the Bellman Equation: From Q-Values to World Models

DGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

safetyarxiv-cs-lg
23 Jun 2026
Safety

IRumAI: Reinforcement Learning for Indian Rummy

DGX agent

arXiv:2606.21975v1 Announce Type: cross Abstract: Despite its massive player base and complex hidden-information dynamics, Indian Rummy has received no reinforcement learning attention. Existing agent

safetyarxiv-cs-lg
23 Jun 2026
Safety

Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes

DGX agent

arXiv:2603.05916v2 Announce Type: replace Abstract: Obstacle avoidance of polytopic obstacles by polytopic robots is a challenging problem in optimization-based control and trajectory planning. Many e

safetyarxiv-cs-ro
23 Jun 2026
Safety

IViT: A Novel Interpretable Visual Transformer for Skin Disease Detection

DGX agent

arXiv:2606.22892v1 Announce Type: cross Abstract: The clinical diagnosis of skin diseases is susceptible to interference from inter-class similarity of skin lesions, and over-reliance on clinicians'ex

safetyarxiv-cs-cv
23 Jun 2026
Safety

JPPD: Joint Prediction_Planning Diffusion with Differentiable Safety Guidance for Dynamic Obstacle Avoidance in Intelligent Transportation Systems

DGX agent

arXiv:2606.20686v1 Announce Type: new Abstract: Shared-space transportation operation requires low-speed autonomous platforms to navigate safely and efficiently among pedestrians, service robots, micr

safetyarxiv-cs-ro
23 Jun 2026
Safety

KITE: Decoupling Kinematics and Interaction for Zero-Shot Cross-Embodiment Manipulation

DGX agent

arXiv:2606.22113v1 Announce Type: new Abstract: Generalizing manipulation policies across robot embodiments remains difficult because standard policies entangle task reasoning with embodiment-specific

safetyarxiv-cs-ro
23 Jun 2026
Safety

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation

DGX agent

arXiv:2606.23685v1 Announce Type: new Abstract: Human-hand demonstrations provide a direct and scalable source of physical interaction data for robot learning. While manual retargeting is indispensabl

safetyarxiv-cs-ro
23 Jun 2026
Safety

Learned Controllers for Agile Quadrotors in Pursuit-Evasion Games

DGX agent

arXiv:2506.02849v3 Announce Type: replace-cross Abstract: In this letter we study 1v1 quadrotor pursuit-evasion, where a pursuer and an evader are trained via reinforcement learning (RL) by competing

safetyarxiv-cs-lg
23 Jun 2026
Safety

Learning Control as Enabling Layer for Embodied Intelligence Research explored with Soft Robotic Swimming in diverse Flow Speeds

DGX agent

arXiv:2606.20660v1 Announce Type: new Abstract: Soft robots are valuable robophysical platforms for studying body-caudal undulatory locomotion, but their compliant bodies are difficult to control prec

safetyarxiv-cs-ro
23 Jun 2026
Safety

Learning Process Rewards via Success Visitation Matching for Efficient RL

DGX agent

arXiv:2606.23640v1 Announce Type: new Abstract: In many modern applications of reinforcement learning (RL), the natural reward for a task of interest is inherently sparse: a reward of 0 is given every

safetyarxiv-cs-lg
23 Jun 2026
Safety

Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem

DGX agent

arXiv:2606.21604v1 Announce Type: new Abstract: Neural combinatorial optimization (NCO) has shown that policies trained by reinforcement can construct strong solutions to NP-hard problems directly fro

safetyarxiv-cs-lg
23 Jun 2026
Safety

Learning to See While Learning to Act: Diffusion Models for Active Perception in Robot Imitation

DGX agent

arXiv:2606.23625v1 Announce Type: new Abstract: Most imitation learning methods assume full observability in table-top settings. In practice, objects are often occluded, requiring robots to both searc

safetyarxiv-cs-ro
23 Jun 2026
Safety

LOLLA: Deep Reinforcement Learning for Closed-Loop Link Adaptation Towards a GPU-Accelerated AI-RAN

DGX agent

arXiv:2606.23110v1 Announce Type: cross Abstract: Outer-loop link adaptation (OLLA) is widely deployed in 5G NR to track channel variations, yet its reliance on first-order, single-bit feedback degrad

safetyarxiv-cs-lg
23 Jun 2026
← Previous
1…8687888990…267
Next →