AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
23 Jun 2026

Improving Reasoning in Vision-Language Models via Perception Verified Self-Training

SafetyDGX agent

arXiv:2606.22158v1 Announce Type: new Abstract: Achieving human-like reasoning in Vision-Language Models (VLMs) remains a long-standing challenge. Recent approaches leverage Chain-of-Thought (CoT) rat

Improving Robotic Imitation Learning via Trajectory Standardization

SafetyDGX agent

arXiv:2606.22907v1 Announce Type: cross Abstract: Imitation learning for robotic manipulation relies on large sets of human demonstration trajectories, which are often noisy and temporally irregular d

Inductive Generalization for Robotic Manipulation

SafetyDGX agent

arXiv:2606.20999v1 Announce Type: cross Abstract: Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models l


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Influencer Cartels

SafetyDGX agent

arXiv:2405.10231v3 Announce Type: replace-cross Abstract: Social media influencers account for a growing share of marketing worldwide. We demonstrate the existence of a novel form of market failure in

'Instead of allowing for greater focus, the latest AI tools are overwhelming workers, frazzling minds and shredding attention spans.' https:…

SafetyDGX agent

'Instead of allowing for greater focus, the latest AI tools are overwhelming workers, frazzling minds and shredding attention spans.' https://www.theatlantic.com/technology/2026/06/ai-agents-jobs-exha

Integrating Facial Generation into Full-Duplex Spoken Dialogue Systems

SafetyDGX agent

arXiv:2606.21970v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models, such as Moshi, enable natural, low-latency voice conversations. However, they remain limited to the audio modality

Intent-Handover: Grounding Language in Human-Usage Regions for Trustworthy Robot-to-Human Handovers

SafetyDGX agent

arXiv:2503.03579v2 Announce Type: replace-cross Abstract: Spoken instructions in robot-to-human handovers may specify either an object ('the cup') or an intended use ('pour water'); in both cases, suc

Interpretable Probabilistic Medical Image Segmentation via Gaussian Process with Explicit Modelling of Annotation Bias and Variability

SafetyDGX agent

arXiv:2606.23177v1 Announce Type: new Abstract: Deep learning-based medical image segmentation models are trained using annotations that exhibit systematic bias and variability across raters. While pr

Inverting the Bellman Equation: From Q-Values to World Models

SafetyDGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

IRumAI: Reinforcement Learning for Indian Rummy

SafetyDGX agent

arXiv:2606.21975v1 Announce Type: cross Abstract: Despite its massive player base and complex hidden-information dynamics, Indian Rummy has received no reinforcement learning attention. Existing agent

Iterative Convex Optimization with Control Barrier Functions for Obstacle Avoidance among Polytopes

SafetyDGX agent

arXiv:2603.05916v2 Announce Type: replace Abstract: Obstacle avoidance of polytopic obstacles by polytopic robots is a challenging problem in optimization-based control and trajectory planning. Many e

IViT: A Novel Interpretable Visual Transformer for Skin Disease Detection

SafetyDGX agent

arXiv:2606.22892v1 Announce Type: cross Abstract: The clinical diagnosis of skin diseases is susceptible to interference from inter-class similarity of skin lesions, and over-reliance on clinicians'ex

JPPD: Joint Prediction_Planning Diffusion with Differentiable Safety Guidance for Dynamic Obstacle Avoidance in Intelligent Transportation Systems

SafetyDGX agent

arXiv:2606.20686v1 Announce Type: new Abstract: Shared-space transportation operation requires low-speed autonomous platforms to navigate safely and efficiently among pedestrians, service robots, micr

KITE: Decoupling Kinematics and Interaction for Zero-Shot Cross-Embodiment Manipulation

SafetyDGX agent

arXiv:2606.22113v1 Announce Type: new Abstract: Generalizing manipulation policies across robot embodiments remains difficult because standard policies entangle task reasoning with embodiment-specific

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation

SafetyDGX agent

arXiv:2606.23685v1 Announce Type: new Abstract: Human-hand demonstrations provide a direct and scalable source of physical interaction data for robot learning. While manual retargeting is indispensabl

Learned Controllers for Agile Quadrotors in Pursuit-Evasion Games

SafetyDGX agent

arXiv:2506.02849v3 Announce Type: replace-cross Abstract: In this letter we study 1v1 quadrotor pursuit-evasion, where a pursuer and an evader are trained via reinforcement learning (RL) by competing

Learning Control as Enabling Layer for Embodied Intelligence Research explored with Soft Robotic Swimming in diverse Flow Speeds

SafetyDGX agent

arXiv:2606.20660v1 Announce Type: new Abstract: Soft robots are valuable robophysical platforms for studying body-caudal undulatory locomotion, but their compliant bodies are difficult to control prec

Learning Process Rewards via Success Visitation Matching for Efficient RL

SafetyDGX agent

arXiv:2606.23640v1 Announce Type: new Abstract: In many modern applications of reinforcement learning (RL), the natural reward for a task of interest is inherently sparse: a reward of 0 is given every

Learning to Place Guards by Reinforcement: A Geo-Free Neural Policy for the Vertex-Guard Art Gallery Problem

SafetyDGX agent

arXiv:2606.21604v1 Announce Type: new Abstract: Neural combinatorial optimization (NCO) has shown that policies trained by reinforcement can construct strong solutions to NP-hard problems directly fro

Learning to See While Learning to Act: Diffusion Models for Active Perception in Robot Imitation

SafetyDGX agent

arXiv:2606.23625v1 Announce Type: new Abstract: Most imitation learning methods assume full observability in table-top settings. In practice, objects are often occluded, requiring robots to both searc

LOLLA: Deep Reinforcement Learning for Closed-Loop Link Adaptation Towards a GPU-Accelerated AI-RAN

SafetyDGX agent

arXiv:2606.23110v1 Announce Type: cross Abstract: Outer-loop link adaptation (OLLA) is widely deployed in 5G NR to track channel variations, yet its reliance on first-order, single-bit feedback degrad

Long-Distance Real-World Navigation of the Legged-Wheeled Robot Go2-W Using Deep Reinforcement Learning

SafetyDGX agent

arXiv:2606.21387v1 Announce Type: new Abstract: Legged-wheeled robots have long been studied for their potential to combine the efficient flat-ground mobility of wheels with the rough-terrain capabili

LP-NavOA: Integrated Local Navigation and Obstacle Avoidance for Humanoid Robots under Limited Perception

SafetyDGX agent

arXiv:2606.23249v1 Announce Type: new Abstract: Humanoid local navigation in cluttered environments must jointly resolve obstacle avoidance, sparse-goal recovery, and stable whole-body locomotion unde

MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning

SafetyDGX agent

arXiv:2606.20641v1 Announce Type: cross Abstract: Multi-modal Large Language Models (MLLMs) have demonstrated remarkable capabilities in semantic understanding and common sense reasoning, making them

Maintain Plasticity in Long-timescale Continual Test-time Adaptation

SafetyDGX agent

arXiv:2412.20034v2 Announce Type: replace Abstract: Continual test-time domain adaptation (CTTA) aims to adjust pre-trained source models to perform well over time across non-stationary target environ

MAPS: Multi-Anchor Projection Similarity for Joint Vision-Language Geo-Localization

SafetyDGX agent

arXiv:2606.22543v1 Announce Type: new Abstract: Humans localize places by integrating perceptual cues from vision with semantic reasoning from language, forming a scene understanding that is both intu

Measuring Model-Induced Discrimination via Efficient Fairness Approximation

SafetyDGX agent

arXiv:2405.09251v2 Announce Type: replace Abstract: Providing various machine learning (ML) applications in the real world, concerns about discrimination hidden in ML models are growing, particularly

Memory Contagion: Cross-Temporal Propagation of Evaluator Bias via Agent Memory

SafetyDGX agent

arXiv:2606.23195v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly rely on memory systems to maintain long-term coherence. Recent work shows that agent memories degrade dur

MemoryVAM: Integrating Memory into Video Action Model for Robot Manipulation

SafetyDGX agent

arXiv:2606.20679v1 Announce Type: cross Abstract: Video-world-model policies learn action-relevant representations by predicting future observations. However, they condition on only a short observatio

Meta copied slot machines to addict kids to Instagram. Now Zuckerberg is turning his company into a prediction market. Meta’s business model…

SafetyDGX agent

Meta copied slot machines to addict kids to Instagram. Now Zuckerberg is turning his company into a prediction market. Meta’s business model is profiting from addiction—kids, gamblers, & more. Stop it

Meta-Reinforcement Learning via Evolution for Multi-Objective Combinatorial Supply Chain Optimisation

SafetyDGX agent

arXiv:2606.22146v1 Announce Type: new Abstract: Meta-reinforcement learning is a promising approach to multi-objective optimisation because it enables rapid policy adaptation across changing environme

Mind the Privileged-to-Camera Gap: Actor-Centric Sidecar Supervision for Camera-First Open-Loop Waypoint Prediction

SafetyDGX agent

arXiv:2606.20772v1 Announce Type: new Abstract: Camera-first autonomous-driving models predict future ego waypoints from images, ego-state features, and route commands, but waypoint supervision alone

Model-Free Robust Average-Reward Reinforcement Learning with Sample Complexity Analysis

SafetyDGX agent

arXiv:2505.12462v3 Announce Type: replace Abstract: Robust reinforcement learning (RL) under the average-reward criterion is essential for long-term decision-making, particularly when the environment

Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning

SafetyDGX agent

arXiv:2606.21943v1 Announce Type: new Abstract: Reinforcement learning (RL) has become central to LLM post-training, yet the methods that dominate current pipelines, PPO and GRPO, represent only a nar

Motion-Aware Reinforcement Learning For Object Localization

SafetyDGX agent

arXiv:2606.21764v1 Announce Type: new Abstract: We present MARLNet (Motion-Aware Reinforcement Learning Network), a PPO-based bounding-box refinement agent that incorporates a constant-velocity motion

MotionPyramid: Hierarchical Motion Representation and Residual Interfaces

SafetyDGX agent

arXiv:2606.20705v1 Announce Type: new Abstract: We ask whether the representational hierarchy seen in perception, from local primitives such as edges to higher level structures such as parts and objec

Multi-Year-to-Decadal Temperature Prediction using a Machine Learning Model-Analog Framework

SafetyDGX agent

arXiv:2502.17583v2 Announce Type: replace-cross Abstract: Multi-year-to-decadal climate predictions are a key tool in understanding the range of potential regional climate futures. Here, we present a

MV-WAM: Manifold-Aware World Action Model with Value Augmentation

SafetyDGX agent

arXiv:2606.21088v1 Announce Type: new Abstract: Achieving robust and generalizable manipulation across diverse environments remains a fundamental challenge in embodied robotics. Recent world action mo

My biggest NY-12 competitors aren't on the ballot today: Trump's AI oligarchs. If they can defeat me—the author of the nation's strongest AI…

SafetyDGX agent

My biggest NY-12 competitors aren't on the ballot today: Trump's AI oligarchs. If they can defeat me—the author of the nation's strongest AI safety bill in one of the nation's bluest districts—they ca

NegAS: Negative Label Guided Attention and Scoring for Out-of-Distribution Object Detection with Vision-Language Models

SafetyDGX agent

arXiv:2606.22537v1 Announce Type: new Abstract: Out-of-Distribution (OOD) detection is essential for ensuring the robustness and reliability of object detection systems deployed in safety-critical app

Neural Architecture Distributions: A New Paradigm for Stochastic Segmentation

SafetyDGX agent

arXiv:2606.21061v1 Announce Type: new Abstract: Stochastic segmentation seeks to represent multiple plausible masks for a single image, which is essential in safety- and quality-critical applications

Neural Conjugate Aggregation: Identifiable Unsupervised Multi-Sensor Regression under Heterogeneous Sensor Bias

SafetyDGX agent

arXiv:2606.22200v1 Announce Type: new Abstract: We study regression-based data fusion under uncertainty, where multiple noisy and biased measurement sources are available but ground-truth labels are a

Noise is Signal: Density-Based Outliers as Leading Indicators of Occupational Emergence in Labor Market Text

SafetyDGX agent

arXiv:2606.22769v1 Announce Type: new Abstract: Standard NLP pipelines for occupational clustering discard the 10-15% of job postings that density-based methods assign to noise. We argue this is an er

Objective-Behavior Alignment: Diagnostics for MORL Policy Selection

SafetyDGX agent

arXiv:2606.21321v1 Announce Type: new Abstract: Real-world decision-making often requires optimizing multiple competing objectives simultaneously. In reinforcement learning (RL), this is typically add

OFMU: Optimization-Driven Framework for Machine Unlearning

SafetyDGX agent

arXiv:2509.22483v2 Announce Type: replace Abstract: Large language models deployed in sensitive applications increasingly require the ability to unlearn specific knowledge, such as user requests, copy

OmniNWM: Omniscient Driving Navigation World Models

SafetyDGX agent

arXiv:2510.18313v5 Announce Type: replace Abstract: Autonomous driving world models are expected to work effectively across three core dimensions: state, action, and reward. However, existing methods

On the Limits of Prompt-Conditioned Language Models as General-Purpose Learners

SafetyDGX agent

arXiv:2606.23668v1 Announce Type: new Abstract: Large Language Models (LLMs) are frequently portrayed as general-purpose solvers capable of solving arbitrary tasks. We argue that this view overlooks a

On the Position Bias of On-Policy Distillation

SafetyDGX agent

arXiv:2606.22600v1 Announce Type: new Abstract: On-Policy Distillation (OPD) improves the learning efficiency of standard reinforcement learning through dense, token-level supervision from teachers. I

One Image is All You Need: Agentic One-Shot Image Generation via Text-Based World Models for Long-Tail Spatial Perception

SafetyDGX agent

arXiv:2606.20764v1 Announce Type: new Abstract: Reliable spatial decision automation, such as autonomous driving and maritime surveillance, critically depends on robust visual perception. However, rea

One Size does not Fit All: Heterogeneous Latent Space Alignment for Unsupervised Domain Adaptation

SafetyDGX agent

arXiv:2606.21415v1 Announce Type: new Abstract: Domain shift remains a major obstacle to the reliable deployment of machine learning models in high-stakes environments such as healthcare. While Domain

oops

SafetyDGX agent

oops Goldman reckons that AI will create about 10trn of discounted value for the world, or up to 22trn. But markets have priced in 27trn of additional value creation. So the size of the 'bubble' has n

OpenHLM: An Empirical Recipe for Whole-Body Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.22174v1 Announce Type: new Abstract: Whole-body humanoid loco-manipulation requires coordinating the robot's entire kinematic chain. However, most existing systems typically decouple the up

'Oracle is under financial pressure because of an expensive build-out of AI data centers for customers like OpenAI.' https://www.bloomberg.c…

SafetyDGX agent

'Oracle is under financial pressure because of an expensive build-out of AI data centers for customers like OpenAI.' https://www.bloomberg.com/news/articles/2026-06-22/oracle-layoffs-fueled-by-ai-redu

Overcoming Imperfect Kinematics in Surgical Robotics Through Sim-to-Real Visuomotor Learning

SafetyDGX agent

arXiv:2606.21396v1 Announce Type: new Abstract: Robot-Assisted Surgery is integral to modern minimally invasive procedures, with automation emerging as the next frontier to enhance precision and reduc

PAIWorld: A 3D-Consistent World Foundation Model for Robotic Manipulation

SafetyDGX agent

arXiv:2606.18375v2 Announce Type: replace Abstract: World foundation models (WFMs) are powerful simulators, yet they predominantly operate in a single-view setting and lack the multi-view 3D consisten

PanoVine: Whole-Body Visuomotor Control for Soft Growing Vine Robot

SafetyDGX agent

arXiv:2606.22923v1 Announce Type: new Abstract: Vine robots, a class of soft, growing robots, are suitable for navigating complex and confined environments due to their compliant bodies and self-suppo

PG-MAP: Joint MAP Optimization for Inference-Time Alignment of Diffusion and Flow-Matching Models

SafetyDGX agent

arXiv:2606.22958v1 Announce Type: cross Abstract: Inference-time alignment of pretrained text-to-image models is typically performed along a single control axis, such as classifier-free guidance, atte

phi-Scene: Physically Grounded Image-to-3D Scene Reconstruction

SafetyDGX agent

arXiv:2606.21596v1 Announce Type: new Abstract: Reconstructing compositional 3D scenes from a single image is a fundamental challenge in 3D world modeling. Recent methods can recover high-fidelity, co

Physically-guided Image Generation for Multi-Projection Mapping

SafetyDGX agent

arXiv:2606.22477v1 Announce Type: new Abstract: Projection Mapping (PM) enables seamless superimposition of digital content onto real-world 3D objects, serving as a fundamental technique for immersive

PISCES: Annotation-free Text-to-Video Post-Training via Optimal Transport-Aligned Rewards

SafetyDGX agent

arXiv:2602.01624v2 Announce Type: replace Abstract: Text-to-video (T2V) generation aims to synthesize videos with high visual quality and temporal consistency that are semantically aligned with input

← Previous
1…6970717273…214
Next →