AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

Meta-Reinforcement Learning via Evolution for Multi-Objective Combinatorial Supply Chain Optimisation

DGX agent

arXiv:2606.22146v1 Announce Type: new Abstract: Meta-reinforcement learning is a promising approach to multi-objective optimisation because it enables rapid policy adaptation across changing environme

safetyarxiv-cs-lg
23 Jun 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Mind the Privileged-to-Camera Gap: Actor-Centric Sidecar Supervision for Camera-First Open-Loop Waypoint Prediction

DGX agent

arXiv:2606.20772v1 Announce Type: new Abstract: Camera-first autonomous-driving models predict future ego waypoints from images, ego-state features, and route commands, but waypoint supervision alone

safetyarxiv-cs-ro
23 Jun 2026
Safety

Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

DGX agent

arXiv:2505.12462v3 Announce Type: replace Abstract: Robust reinforcement learning (RL) under the average-reward criterion is essential for long-term decision-making, particularly when the environment

safetyarxiv-cs-lg
23 Jun 2026
Safety

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning

DGX agent

arXiv:2606.21943v1 Announce Type: new Abstract: Reinforcement learning (RL) has become central to LLM post-training, yet the methods that dominate current pipelines, PPO and GRPO, represent only a nar

safetyarxiv-cs-lg
23 Jun 2026
Safety

Motion-Aware Reinforcement Learning For Object Localization

DGX agent

arXiv:2606.21764v1 Announce Type: new Abstract: We present MARLNet (Motion-Aware Reinforcement Learning Network), a PPO-based bounding-box refinement agent that incorporates a constant-velocity motion

safetyarxiv-cs-cv
23 Jun 2026
Safety

MotionPyramid: Hierarchical Motion Representation and Residual Interfaces

DGX agent

arXiv:2606.20705v1 Announce Type: new Abstract: We ask whether the representational hierarchy seen in perception, from local primitives such as edges to higher level structures such as parts and objec

safetyarxiv-cs-cv
23 Jun 2026
Safety

Multi-Year-to-Decadal Temperature Prediction using a Machine Learning Model-Analog Framework

DGX agent

arXiv:2502.17583v2 Announce Type: replace-cross Abstract: Multi-year-to-decadal climate predictions are a key tool in understanding the range of potential regional climate futures. Here, we present a

safetyarxiv-cs-lg
23 Jun 2026
Safety

MV-WAM: Manifold-Aware World Action Model with Value Augmentation

DGX agent

arXiv:2606.21088v1 Announce Type: new Abstract: Achieving robust and generalizable manipulation across diverse environments remains a fundamental challenge in embodied robotics. Recent world action mo

safetyarxiv-cs-ro
23 Jun 2026
Safety

Neural Conjugate Aggregation: Identifiable Unsupervised Multi-Sensor Regression under Heterogeneous Sensor Bias

DGX agent

arXiv:2606.22200v1 Announce Type: new Abstract: We study regression-based data fusion under uncertainty, where multiple noisy and biased measurement sources are available but ground-truth labels are a

safetyarxiv-cs-lg
23 Jun 2026
Safety

Objective-Behavior Alignment: Diagnostics for MORL Policy Selection

DGX agent

arXiv:2606.21321v1 Announce Type: new Abstract: Real-world decision-making often requires optimizing multiple competing objectives simultaneously. In reinforcement learning (RL), this is typically add

safetyarxiv-cs-lg
23 Jun 2026
Safety

OmniNWM: Omniscient Driving Navigation World Models

DGX agent

arXiv:2510.18313v5 Announce Type: replace Abstract: Autonomous driving world models are expected to work effectively across three core dimensions: state, action, and reward. However, existing methods

safetyarxiv-cs-cv
23 Jun 2026
Safety

On the Position Bias of On-Policy Distillation

DGX agent

arXiv:2606.22600v1 Announce Type: new Abstract: On-Policy Distillation (OPD) improves the learning efficiency of standard reinforcement learning through dense, token-level supervision from teachers. I

safetyarxiv-cs-lg
23 Jun 2026
Safety

One Size does not Fit All: Heterogeneous Latent Space Alignment for Unsupervised Domain Adaptation

DGX agent

arXiv:2606.21415v1 Announce Type: new Abstract: Domain shift remains a major obstacle to the reliable deployment of machine learning models in high-stakes environments such as healthcare. While Domain

safetyarxiv-cs-lg
23 Jun 2026
Safety

oops

DGX agent

oops Goldman reckons that AI will create about 10trn of discounted value for the world, or up to 22trn. But markets have priced in 27trn of additional value creation. So the size of the 'bubble' has n

safetygary-marcus--x
23 Jun 2026
Safety

OpenHLM: An Empirical Recipe for Whole-Body Humanoid Loco-Manipulation

DGX agent

arXiv:2606.22174v1 Announce Type: new Abstract: Whole-body humanoid loco-manipulation requires coordinating the robot's entire kinematic chain. However, most existing systems typically decouple the up

safetyarxiv-cs-ro
23 Jun 2026
Safety

'Oracle is under financial pressure because of an expensive build-out of AI data centers for customers like OpenAI.' https://www.bloomberg.c…

DGX agent

'Oracle is under financial pressure because of an expensive build-out of AI data centers for customers like OpenAI.' https://www.bloomberg.com/news/articles/2026-06-22/oracle-layoffs-fueled-by-ai-redu

safetygary-marcus--x
23 Jun 2026
Safety

Overcoming Imperfect Kinematics in Surgical Robotics Through Sim-to-Real Visuomotor Learning

DGX agent

arXiv:2606.21396v1 Announce Type: new Abstract: Robot-Assisted Surgery is integral to modern minimally invasive procedures, with automation emerging as the next frontier to enhance precision and reduc

safetyarxiv-cs-ro
23 Jun 2026
Safety

PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

DGX agent

arXiv:2606.18375v2 Announce Type: replace Abstract: World foundation models (WFMs) are powerful simulators, yet they predominantly operate in a single-view setting and lack the multi-view 3D consisten

safetyarxiv-cs-ro
23 Jun 2026
Safety

PanoVine: Whole-Body Visuomotor Control for Soft Growing Vine Robot

DGX agent

arXiv:2606.22923v1 Announce Type: new Abstract: Vine robots, a class of soft, growing robots, are suitable for navigating complex and confined environments due to their compliant bodies and self-suppo

safetyarxiv-cs-ro
23 Jun 2026
Safety

PG-MAP: Joint MAP Optimization for Inference-Time Alignment of Diffusion and Flow-Matching Models

DGX agent

arXiv:2606.22958v1 Announce Type: cross Abstract: Inference-time alignment of pretrained text-to-image models is typically performed along a single control axis, such as classifier-free guidance, atte

safetyarxiv-cs-cv
23 Jun 2026
Safety

phi-Scene: Physically Grounded Image-to-3D Scene Reconstruction

DGX agent

arXiv:2606.21596v1 Announce Type: new Abstract: Reconstructing compositional 3D scenes from a single image is a fundamental challenge in 3D world modeling. Recent methods can recover high-fidelity, co

safetyarxiv-cs-cv
23 Jun 2026
Safety

Physically-guided Image Generation for Multi-Projection Mapping

DGX agent

arXiv:2606.22477v1 Announce Type: new Abstract: Projection Mapping (PM) enables seamless superimposition of digital content onto real-world 3D objects, serving as a fundamental technique for immersive

safetyarxiv-cs-cv
23 Jun 2026
Safety

PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards

DGX agent

arXiv:2602.01624v2 Announce Type: replace Abstract: Text-to-video (T2V) generation aims to synthesize videos with high visual quality and temporal consistency that are semantically aligned with input

safetyarxiv-cs-cv
23 Jun 2026
Safety

PJ-RoPE: A Fourier-Jet-Affine Position Space for Relative Attention

DGX agent

arXiv:2606.05345v2 Announce Type: replace Abstract: We organize relative-position mechanisms in attention as a learnable Fourier-Jet-Affine position space. The starting point is lag-shift dynamics: a

safetyarxiv-cs-lg
23 Jun 2026
Safety

PoLAR: Factorizing Extent and Mode in Latent Actions for Robot Policy Learning

DGX agent

arXiv:2606.21139v1 Announce Type: cross Abstract: Latent action pretraining learns representations of visual change from pairs of observations, but existing methods typically encode each transition as

safetyarxiv-cs-lg
23 Jun 2026
Safety

Policy-as-Data: Learning Generalizable HOI Diffusion Models from Simulated Physics

DGX agent

arXiv:2606.22806v1 Announce Type: new Abstract: Synthesizing realistic Human-Object Interactions (HOI) is critical for creating embodied avatars and functional virtual environments. However, current d

safetyarxiv-cs-cv
23 Jun 2026
Safety

PolicyTrim: Boosting Intrinsic Policy Efficiency of Vision-Language-Action Models

DGX agent

arXiv:2606.22540v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models provide a unified paradigm for robotic manipulation, yet their real-world deployment is often bottlenecked by execut

safetyarxiv-cs-cv
23 Jun 2026
Safety

Pose-Agnostic Robotic Functional Grasping via Observation-Action Canonicalization

DGX agent

arXiv:2606.21148v1 Announce Type: new Abstract: Functional robotic grasping requires a policy that generalizes across diverse object geometries and poses while maintaining task-specific contact precis

safetyarxiv-cs-ro
23 Jun 2026
Safety

Pose Anything Anywhere:Model-free Object Poses from Arbitrary References

DGX agent

arXiv:2606.23634v1 Announce Type: new Abstract: Estimating the 6D pose of unseen objects is a fundamental yet challenging problem for open-world robotics and embodied perception. Model-based methods a

safetyarxiv-cs-cv
23 Jun 2026
Safety

Pose6DAug: Physically Plausible Multi-view Object Swapping for Robot Data Augmentation

DGX agent

arXiv:2606.20118v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) policies have shown strong potential for general-purpose manipulation, yet they often fail on novel, out-of-distr

safetyarxiv-cs-lg
23 Jun 2026
Safety

Prefix-Guided On-Policy Distillation: Mining Golden Trajectories from Rollouts

DGX agent

arXiv:2606.21994v1 Announce Type: new Abstract: On-policy distillation (OPD) improves reasoning models by applying dense teacher supervision on student-sampled trajectories. However, scaling OPD to lo

safetyarxiv-cs-lg
23 Jun 2026
Safety

Programmable magnetic soft robots with controlled locomotion and directional liquid cargo release

DGX agent

arXiv:2606.21737v1 Announce Type: new Abstract: Magnetically programmable soft elastomers enable complex shape morphing and locomotion dynamics in small scale soft robots under external magnetic field

safetyarxiv-cs-ro
23 Jun 2026
Safety

Provably Efficient Policy-Reward Co-Pretraining for Adversarial Imitation Learning

DGX agent

arXiv:2606.22056v1 Announce Type: new Abstract: Adversarial imitation learning (AIL) achieves high-quality imitation compared to behavioral cloning (BC), but demands substantial online environment int

safetyarxiv-cs-lg
23 Jun 2026
Safety

RECALL: Recovery Experience Collection for Active Lifelong Learning in Vision-Language-Action Models

DGX agent

arXiv:2606.23617v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are commonly fine-tuned through passive imitation learning, where additional demonstrations are collected for task

safetyarxiv-cs-lg
23 Jun 2026
Safety

Recency/Frequency Adaptive KV Caching for Large Language Model Serving

DGX agent

arXiv:2606.21238v1 Announce Type: cross Abstract: Key-value (KV) caching is a powerful technique for accelerating large language model inference and generation. Inference workloads are large and diver

safetyarxiv-cs-lg
23 Jun 2026
Safety

ReFPO: Reflow Regularization for Flow Matching Policy Gradients

DGX agent

arXiv:2606.21086v1 Announce Type: new Abstract: We present Reflow-regularized Flow Matching Policy Gradients (ReFPO), a simple online RL method that adds explicit Reflow regularization to FPO for effi

safetyarxiv-cs-ro
23 Jun 2026
Safety

Reinforcement Learning-Based Traffic Signal Control for IoT-Enabled Intersections

DGX agent

arXiv:2606.22108v1 Announce Type: cross Abstract: Urban traffic congestion remains a persistent challenge in car-dependent cities, imposing significant economic and societal costs. Traffic signal syst

safetyarxiv-cs-lg
23 Jun 2026
Safety

RelightAnyone: A Generalized Relightable 3D Gaussian Head Model

DGX agent

arXiv:2601.03357v2 Announce Type: replace Abstract: 3D Gaussian Splatting (3DGS) has become a standard approach to reconstruct and render photorealistic 3D head avatars. A major challenge is to religh

safetyarxiv-cs-cv
23 Jun 2026
Safety

Remember what you did?: Learning Behavioral Memories for Partially Observable Object Manipulation

DGX agent

arXiv:2606.21188v1 Announce Type: new Abstract: Long horizon, contact-rich manipulation is inherently partially observable. This is as a single visual observation rarely captures a robot's full action

safetyarxiv-cs-ro
23 Jun 2026
Model Releases

Residue-Level Attributions in Protein Language Models Do Not Recover Allergen Epitopes

DGX agent

arXiv:2606.22181v1 Announce Type: new Abstract: Deep allergenicity classifiers are increasingly used in safety screening of novel foods, and recent protein language models have substantially improved

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Rethinking Object-Centric Representations for Video Dynamics Modeling

DGX agent

arXiv:2606.23436v1 Announce Type: new Abstract: Unsupervised video object tracking aims to decompose dynamic scenes into persistent, object-centric entities without manual annotations. Many recent app

safetyarxiv-cs-cv
23 Jun 2026
Safety

RoboLineage: Agent-Native Data Lifecycle Governance Across Robot Policy Iterations

DGX agent

arXiv:2606.22142v1 Announce Type: new Abstract: We present RoboLineage, an agent-native data lifecycle governance system for robot policy iteration. Modern robot policies improve through repeated data

safetyarxiv-cs-ro
23 Jun 2026
Safety

Robot Critics that Sweat the Small Stuff

DGX agent

arXiv:2606.21572v1 Announce Type: new Abstract: Large vision-language models contain several priors about the world and object interactions, making them useful critics during inference to steer robot

safetyarxiv-cs-ro
23 Jun 2026
Safety

Robot Self-Improvement via Human-Video Dynamics Models

DGX agent

arXiv:2606.21406v1 Announce Type: cross Abstract: A central question in robot learning is how to acquire skills from the kinds of data that humans learn from: passive observation, embodied practice, a

safetyarxiv-cs-cv
23 Jun 2026
Safety

Robust Representation Learning in Masked Autoencoders

DGX agent

arXiv:2602.03531v2 Announce Type: replace-cross Abstract: Masked Autoencoders (MAEs) achieve impressive performance in image classification tasks, yet the internal representations they learn remain le

safetyarxiv-cs-cv
23 Jun 2026
Safety

Rotation-Aware Point-Cloud Embeddings for Vision-Based In-Hand Reorientation

DGX agent

arXiv:2606.21788v1 Announce Type: cross Abstract: Point-cloud goals provide a direct way to specify dexterous in-hand reorientation: instead of defining an object-specific pose frame or estimating 6D

safetyarxiv-cs-cv
23 Jun 2026
Safety

RubricRL: Simple Generalizable Rewards for Text-to-Image Generation

DGX agent

arXiv:2511.20651v2 Announce Type: replace Abstract: Reinforcement learning (RL) has recently emerged as a promising approach for aligning text-to-image generative models with human preferences. A key

safetyarxiv-cs-cv
23 Jun 2026
Safety

SAGE: An Expert-Annotated South Asian GI Endoscopy Dataset for Multimodal Learning and Hallucination Analysis

DGX agent

arXiv:2606.22144v1 Announce Type: new Abstract: Gastrointestinal cancers represent a growing health burden in the South Asian region, driven largely by rapid changes in socio-economic conditions & lif

safetyarxiv-cs-cv
23 Jun 2026
← Previous
1…134135136137138…302
Next →