AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-Training

DGX agent

arXiv:2607.04574v1 Announce Type: cross Abstract: For LLM agents, supervised fine-tuning is not only about teacher labels' quality, but also about which interaction contexts those labels condition on.

safetyarxiv-cs-ai
7 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

A Graph-Based Reinforcement Learning Approach with Frontier Potential Based Reward for Safe Cluttered Environment Exploration

DGX agent

arXiv:2504.11907v3 Announce Type: replace Abstract: Autonomous exploration of cluttered environments requires efficient exploration strategies that guarantee safety against potential collisions with u

safetyarxiv-cs-ro
7 Jul 2026
Safety

A Hierarchy of Policy Learning Problems

DGX agent

arXiv:2607.03385v1 Announce Type: cross Abstract: Policy learning has received substantial attention with the goal of learning policies from observational data for decision-making. A majority of work

safetyarxiv-cs-lg
7 Jul 2026
Safety

A Mathematical Theory of Value: a synthesis on goal-directed agency under resource constraints

DGX agent

arXiv:2606.12502v2 Announce Type: replace-cross Abstract: We propose that value -- the quantity goal-directed agents create, destroy, and exchange -- is a lawful structural quantity in the same catego

safetyarxiv-cs-ai
7 Jul 2026
Safety

A Perception-Manipulation Robotics System for Food Cutting

DGX agent

arXiv:2607.04367v1 Announce Type: new Abstract: In the development of cooking robots, mastering the task of cutting is crucial. A significant challenge lies in the diverse properties of food, which ne

safetyarxiv-cs-ro
7 Jul 2026
Safety

A Policy Decomposition Framework for Dynamic Order Fulfillment Operations

DGX agent

arXiv:2607.04056v1 Announce Type: cross Abstract: Modern supply chains span diverse operational environments, ranging from e-commerce distribution networks to customized production-to-order manufactur

safetyarxiv-cs-lg
7 Jul 2026
Safety

A Precedent-Guided Co-Scientist for Side-Effect-Aware Drug Redesign

DGX agent

arXiv:2607.02944v1 Announce Type: cross Abstract: We propose PRECEDE, a precedent-guided co-scientist for side-effect-aware drug redesign that revises a parent compound to mitigate a specified side ef

safetyarxiv-cs-ai
7 Jul 2026
Safety

A Unified Causal-Origin Taxonomy of Distributional Shifts in Reinforcement Learning

DGX agent

arXiv:2606.16933v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) systems often degrade when operating conditions differ from those previously encountered, reflecting distributiona

safetyarxiv-cs-ai
7 Jul 2026
Safety

A User-driven Design Framework for Robotaxi

DGX agent

arXiv:2602.19107v3 Announce Type: replace Abstract: Robotaxis are emerging as a promising form of urban mobility, but removing human drivers fundamentally reshapes passenger-vehicle interaction and ra

safetyarxiv-cs-ro
7 Jul 2026
Safety

ACE: Agentic Control for Embodied Manipulation via Zero-shot Workflow Reasoning

DGX agent

arXiv:2607.04162v1 Announce Type: cross Abstract: Open-ended tabletop manipulation requires agents to not only understand natural language but also adapt to dynamic environments and execution failures

safetyarxiv-cs-lg
7 Jul 2026
Safety

ACPO: Adaptive Credit Policy Optimization via Fine-Grained Surrogate Entropy

DGX agent

arXiv:2607.03126v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has substantially improved the reasoning ability of large language models (LLMs), but sparse outcome rewards still make to

safetyarxiv-cs-ai
7 Jul 2026
Safety

Adaptive Entropy-Driven Sensor Selection in a Camera-LiDAR Particle Filter for Single-Vessel Tracking

DGX agent

arXiv:2603.08457v2 Announce Type: replace-cross Abstract: Robust single-vessel tracking from fixed coastal platforms is hindered by modality-specific degradations: cameras suffer from illumination and

safetyarxiv-cs-lg
7 Jul 2026
Safety

Adaptive Inference Batching using Policy Gradients

DGX agent

arXiv:2607.05272v1 Announce Type: cross Abstract: Inference serving systems must balance throughput and latency under bursty, heterogeneous workloads, yet the industry standard remains static batching

safetyarxiv-cs-ai
7 Jul 2026
Safety

Adaptive Margin RLHF via Preference over Preferences

DGX agent

arXiv:2509.22851v4 Announce Type: replace-cross Abstract: Margin-based optimization is fundamental to improving generalization and robustness in classification tasks. In the context of reward model le

safetyarxiv-cs-ai
7 Jul 2026
Safety

Adaptive Partitioning and Learning for Stochastic Control of Diffusion Processes

DGX agent

arXiv:2512.14991v2 Announce Type: replace Abstract: We study reinforcement learning for controlled diffusion processes with unbounded continuous state spaces, bounded continuous actions, and polynomia

safetyarxiv-cs-lg
7 Jul 2026
Safety

Additive Causal Construction for Transferable and Reconfigurable Cross-System Learning in Multi-Source Image Fusion

DGX agent

arXiv:2607.02572v1 Announce Type: cross Abstract: In multi-source image fusion scenarios, heterogeneous inputs are typically driven by distinct generative mechanisms and can be viewed as a composition

safetyarxiv-cs-ai
7 Jul 2026
Safety

ADP: Adversarial Dynamics Priors for Physically Grounded Humanoid Locomotion

DGX agent

arXiv:2607.03454v1 Announce Type: cross Abstract: In this paper, we propose Adversarial Dynamics Priors (ADP) for perturbation-resilient humanoid locomotion control. Existing motion prior-based method

safetyarxiv-cs-lg
7 Jul 2026
Safety

Agentic Artificial Intelligence for Multistage Physics Experiments at a Large-Scale User Facility Particle Accelerator

DGX agent

arXiv:2509.17255v2 Announce Type: replace-cross Abstract: We present the first language-model-driven agentic artificial intelligence (AI) system to autonomously execute multi-stage physics experiments

safetyarxiv-cs-ai
7 Jul 2026
Safety

AGL-1: The Enterprise AI Governance Layer as a Control Plane for Trusted Enterprise Intelligence

DGX agent

arXiv:2607.03516v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from isolated experimentation toward operational dependency across copilots, retrieval-augmented generati

safetyarxiv-cs-ai
7 Jul 2026
Safety

Aligning Language Models with Selective Prediction

DGX agent

arXiv:2607.03528v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as critical decision-making components in high-stakes real-world AI systems, rendering LLM reli

safetyarxiv-cs-ai
7 Jul 2026
Safety

Alignment-Guided Largest Table Overlap Size Estimation

DGX agent

arXiv:2607.03049v1 Announce Type: new Abstract: Fast estimation of the size of the largest overlap between tables enables blocking and query-by-table retrieval in large table repositories. The first a

safetyarxiv-cs-cl
7 Jul 2026
Safety

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis

DGX agent

arXiv:2512.11797v2 Announce Type: replace-cross Abstract: The collection of large-scale and diverse robot demonstrations remains a major bottleneck for imitation learning, as real-world data acquisiti

safetyarxiv-cs-cv
7 Jul 2026
Safety

Anticipatory Reinforcement Learning for Trajectory Tracking

DGX agent

arXiv:2607.03132v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) in industrial control often suffers from lag and overshoot due to purely reactive control based on the current trackin

safetyarxiv-cs-lg
7 Jul 2026
Safety

AquaStereo: Enabling Underwater Stereo Matching via Depth-Conditioned Diffusion and Geometry Self-Distillation

DGX agent

arXiv:2607.04303v1 Announce Type: new Abstract: Learning-based stereo matching models struggle in underwater environments due to scarce in-domain data and the difficulty of extracting discriminative c

safetyarxiv-cs-cv
7 Jul 2026
Safety

ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning

DGX agent

arXiv:2602.21534v3 Announce Type: replace Abstract: Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interacti

safetyarxiv-cs-ai
7 Jul 2026
Safety

As the @UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications fo…

DGX agent

As the @UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications for global collaboration and policymaking. I’m hopeful that we

safetyyoshua-bengio--x
7 Jul 2026
Safety

ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability

DGX agent

arXiv:2607.02686v1 Announce Type: new Abstract: Reinforcement learning agents operating under partial observability must act on incomplete information, making them natural candidates for guidance from

safetyarxiv-cs-ai
7 Jul 2026
Safety

At UMA, we build and own the full stack, from hardware to software. This allows us to bake safety in at all levels, rather than bolt it on a…

DGX agent

At UMA, we build and own the full stack, from hardware to software. This allows us to bake safety in at all levels, rather than bolt it on as an afterthought. Because trust is everything when robots s

safetyyann-lecun--x
7 Jul 2026
Safety

Athena-WBC: Capability-Aligned Policy Experts for Long-Tail Humanoid Whole-Body Control

DGX agent

arXiv:2607.04837v1 Announce Type: new Abstract: Large-scale humanoid motion-tracking controllers are commonly improved by reallocating training effort: difficult motions are sampled more often, isolat

safetyarxiv-cs-ro
7 Jul 2026
Safety

Attention Limited Reward Learning

DGX agent

arXiv:2607.04590v1 Announce Type: new Abstract: Pairwise human comparisons are a primary interface through which modern AI systems learn human preferences. RLHF and related alignment pipelines typical

safetyarxiv-cs-ai
7 Jul 2026
Safety

Attributing Emergence in Million-Agent Systems

DGX agent

arXiv:2605.11404v2 Announce Type: replace Abstract: Large language models (LLMs) can simulate human-like reasoning and decision-making in individual agents. LLM-powered multi-agent systems (MAS) combi

safetyarxiv-cs-ai
7 Jul 2026
Safety

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment

DGX agent

arXiv:2607.04311v1 Announce Type: new Abstract: Subject-driven and multi-element video generation are central to controllable video synthesis, but existing methods still struggle to preserve identity

safetyarxiv-cs-cv
7 Jul 2026
Safety

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

DGX agent

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward aut

safetyarxiv-cs-cl
7 Jul 2026
Safety

AViS-Mamba: Adaptive Visual Steering of Audio State-Space Dynamics for Violence Detection

DGX agent

arXiv:2604.03329v2 Announce Type: replace-cross Abstract: Automatic violence detection from video is challenging because violent interactions may be distant, occluded, or only partially visible. Audio

safetyarxiv-cs-ai
7 Jul 2026
Safety

Benign Overfitting Does Not Occur in Diffusion Models

DGX agent

arXiv:2607.02671v1 Announce Type: cross Abstract: Benign overfitting and double descent have come to shape our understanding of generalization in deep learning, establishing that overfitting is not on

safetyarxiv-cs-lg
7 Jul 2026
Safety

Best-of-Better-N: Generating Pre-Aligned Responses with In-Context Learning

DGX agent

arXiv:2607.03453v1 Announce Type: cross Abstract: Inference-time alignment methods, such as Best-of-N, offer a flexible alternative to training-based alignment by using reward models to select high-qu

safetyarxiv-cs-ai
7 Jul 2026
Safety

BEVLM: Distilling Semantic Knowledge from LLMs into Bird's-Eye View Representations

DGX agent

arXiv:2603.06576v2 Announce Type: replace-cross Abstract: The integration of Large Language Models (LLMs) into autonomous driving has attracted growing interest for their strong reasoning and semantic

safetyarxiv-cs-ai
7 Jul 2026
Safety

Beyond Heuristics: A Standardized Real2Sim Pipeline for Physical Human Robot Interaction in Human-in-the-Loop Simulation

DGX agent

arXiv:2607.03017v1 Announce Type: new Abstract: The aging global population drives demand for assistive robots, yet the safety risks and costs of physical testing make Human-in-the-Loop (HITL) simulat

safetyarxiv-cs-ro
7 Jul 2026
Safety

Beyond Independent Labels: Schwartz-Geometry Decoding for Human Value Detection

DGX agent

arXiv:2607.05052v1 Announce Type: cross Abstract: Human value detection is commonly formulated as sentence-level multi-label classification over the 19 refined Schwartz values, typically predicted as

safetyarxiv-cs-ai
7 Jul 2026
Safety

Beyond Point-Attached Semantics: Object-Centric Semantic Fields for Generalizable Manipulation

DGX agent

arXiv:2607.03163v1 Announce Type: new Abstract: Generalizable robot manipulation requires stable 3D understanding of functional object parts, such as handles, tool heads, openings, and graspable regio

safetyarxiv-cs-ro
7 Jul 2026
Safety

Beyond Random Sampling: Distribution-Aware Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2607.04249v1 Announce Type: new Abstract: Precise medical image segmentation is crucial for clinical diagnosis and treatment planning, yet relies heavily on expensive expert annotations. Semi-su

safetyarxiv-cs-cv
7 Jul 2026
Safety

BGP route policies: Top 3 use cases by customer demand

DGX agent

When we first made BGP route policies for Cloud Router generally available over a year ago, our goal was to give network administrators deep, programmable control over how network paths are evaluated

safetygoogle-cloud-ai
7 Jul 2026
Safety

BiSLW: Bi-Spectral Latent Watermarking for Generative Diffusion Models

DGX agent

arXiv:2607.02643v1 Announce Type: new Abstract: Diffusion-based generative models have transformed visual content synthesis, yet they remain vulnerable to unauthorized usage and lack reliable attribut

safetyarxiv-cs-cv
7 Jul 2026
Safety

Bootstrap Flow-Map Tree Sampling Enables Online Feedback Driven Search

DGX agent

arXiv:2607.02915v1 Announce Type: cross Abstract: In many scientific and engineering domains, maximizing discovery within a limited sampling budget demands strategic, observation-guided exploration. W

safetyarxiv-cs-ai
7 Jul 2026
Safety

BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation

DGX agent

arXiv:2601.18253v2 Announce Type: replace-cross Abstract: Accurate evaluation of user satisfaction is critical for iterative development of conversational AI. However, for open-ended assistants, tradi

safetyarxiv-cs-ai
7 Jul 2026
Safety

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process

DGX agent

arXiv:2607.03748v1 Announce Type: new Abstract: Unified multi-modal models (UMMs) have shown promising interleaved text-image reasoning capabilities, yet effectively optimizing such multi-turn generat

safetyarxiv-cs-ai
7 Jul 2026
Safety

CAGE-1: Control, Assurance, and Governance Evaluation for Enterprise Agentic AI

DGX agent

arXiv:2607.03510v1 Announce Type: cross Abstract: Enterprise artificial intelligence is moving from experimentation into operational workflows. Early programs focused on model access and retrieval-aug

safetyarxiv-cs-ai
7 Jul 2026
Safety

Can temporal article-level credibility signals improve domain-level credibility prediction?

DGX agent

arXiv:2607.04560v1 Announce Type: new Abstract: Web domain credibility evaluation is vital for combating misinformation. It is conducted by examining factors such as domain type, transparency, and ove

safetyarxiv-cs-cl
7 Jul 2026
← Previous
1…5152535455…265
Next →