AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Temper and Tilt Lead to SLOP: Reward Hacking Mitigation with Inference-Time Alignment

DGX agent

arXiv:2605.13537v1 Announce Type: cross Abstract: Inference-time alignment techniques offer a lightweight alternative or complement to costly reinforcement learning, while enabling continual adaptatio

safetyarxiv-cs-ai
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Test-time Offline Reinforcement Learning on Goal-related Experience

DGX agent

arXiv:2507.18809v2 Announce Type: replace Abstract: Foundation models compress a large amount of information in a single, large neural network, which can then be queried for individual tasks. There ar

safetyarxiv-cs-lg
14 May 2026
Safety

Test-time Sparsity for Extreme Fast Action Diffusion

DGX agent

arXiv:2605.13316v1 Announce Type: new Abstract: Action diffusion excels at high-fidelity action generation but incurs heavy computational costs owing to its iterative denoising nature. Despite current

safetyarxiv-cs-cv
14 May 2026
Safety

The End Justifies the Mean: A Linear Ranking Rule for Proportional Sequential Decisions

DGX agent

arXiv:2605.12717v1 Announce Type: cross Abstract: AI alignment and participatory design motivate a new democratic design problem: how to collectively choose a decision rule to use repeatedly. We study

safetyarxiv-cs-ai
14 May 2026
Safety

The Horizon Threshold in Cooperative Multi-Agent Reward-Free Exploration

DGX agent

arXiv:2602.01453v3 Announce Type: replace Abstract: We study cooperative multi-agent reinforcement learning in the setting of reward-free exploration, where multiple agents jointly explore an unknown

safetyarxiv-cs-lg
14 May 2026
Safety

Think Twice, Act Once: Verifier-Guided Action Selection For Embodied Agents

DGX agent

arXiv:2605.12620v1 Announce Type: new Abstract: Building generalist embodied agents capable of solving complex real-world tasks remains a fundamental challenge in AI. Multimodal Large Language Models

safetyarxiv-cs-ai
14 May 2026
Safety

Tight Sample Complexity Bounds for Entropic Best Policy Identification

DGX agent

arXiv:2605.13717v1 Announce Type: new Abstract: We study best-policy identification for finite-horizon risk-sensitive reinforcement learning under the entropic risk measure. Recent work established a

safetyarxiv-cs-lg
14 May 2026
Safety

Topology-Preserving Neural Operator Learning via Hodge Decomposition

DGX agent

arXiv:2605.13834v1 Announce Type: cross Abstract: In this paper, we study solution operators of physical field equations on geometric meshes from a function-space perspective. We reveal that Hodge ort

safetyarxiv-cs-ai
14 May 2026
Safety

Towards a holistic understanding of Selection Bias for Causal Effect Identification

DGX agent

arXiv:2605.13430v1 Announce Type: cross Abstract: Selection bias is pervasive in observational studies. For example, large scale biobanks data can exhibit ``healthy volunteer bias'' when respondents a

safetyarxiv-cs-ai
14 May 2026
Safety

Towards Generalizable Reasoning: Group Causal Counterfactual Policy Optimization for LLM Reasoning

DGX agent

arXiv:2602.06475v2 Announce Type: replace Abstract: Large language models (LLMs) excel at complex tasks with advances in reasoning capabilities. However, existing reward mechanisms remain tightly coup

safetyarxiv-cs-lg
14 May 2026
Safety

TrackCraft3R: Repurposing Video Diffusion Transformers for Dense 3D Tracking

DGX agent

arXiv:2605.12587v1 Announce Type: new Abstract: Dense 3D tracking from monocular video is fundamental to dynamic scene understanding. While recent 3D foundation models provide reliable per-frame geome

safetyarxiv-cs-cv
14 May 2026
Safety

Trajectory-Level Data Augmentation for Offline Reinforcement Learning

DGX agent

arXiv:2605.13401v1 Announce Type: new Abstract: We propose a data augmentation method for offline reinforcement learning, motivated by active positioning problems. Particularly, our approach enables t

safetyarxiv-cs-lg
14 May 2026
Safety

Uncertainty-aware Spatial-Frequency Registration and Fusion for Infrared and Visible Images

DGX agent

arXiv:2605.13049v1 Announce Type: new Abstract: Infrared and Visible Image Fusion (IVIF) has shown promise in visual tasks under challenging environments, but fusion under unregistered conditions face

safetyarxiv-cs-cv
14 May 2026
Safety

Unifying Entropy Regularization in Optimal Control: From and Back to Classical Objectives via Iterated Soft Policies and Path Integral Solutions

DGX agent

arXiv:2512.06109v3 Announce Type: replace-cross Abstract: This paper develops a unified perspective on several optimal control formulations through the lens of Kullback-Leibler (KL) regularization. We

safetyarxiv-cs-lg
14 May 2026
Safety

UniJEPA: Enhancing Robot Policy via Unified Continuous and Discrete Representation Learning

DGX agent

arXiv:2510.10642v3 Announce Type: replace-cross Abstract: Building generalist robot policies that can handle diverse tasks in open-ended environments is a central challenge in robotics. To leverage kn

safetyarxiv-cs-ai
14 May 2026
Safety

Unweighted ranking for value-based decision making with uncertainty

DGX agent

arXiv:2605.13601v1 Announce Type: new Abstract: As intelligent systems are increasingly implemented in our society to make autonomous decisions, their commitment to human values raises serious concern

safetyarxiv-cs-ai
14 May 2026
Safety

VideoSEAL: Mitigating Evidence Misalignment in Agentic Long Video Understanding by Decoupling Answer Authority

DGX agent

arXiv:2605.12571v1 Announce Type: cross Abstract: Long video question answering requires locating sparse, time-scattered visual evidence within highly redundant content. Although current MLLMs perform

safetyarxiv-cs-ai
14 May 2026
Safety

WD-FQDet: Multispectral Detection Transformer via Wavelet Decomposition and Frequency-aware Query Learning

DGX agent

arXiv:2605.13621v1 Announce Type: new Abstract: Infrared-visible object detection improves detection performance by combining complementary features from multispectral images. Existing backbone-specif

safetyarxiv-cs-cv
14 May 2026
Safety

What properties of reasoning supervision are associated with improved downstream model quality?

DGX agent

arXiv:2605.13290v1 Announce Type: new Abstract: Validating training data for reasoning models typically requires expensive trial-and-error fine-tuning cycles. In this work, we investigate whether the

safetyarxiv-cs-ai
14 May 2026
Safety

What to Ignore, What to React: Visually Robust RL Fine-Tuning of VLA Models

DGX agent

arXiv:2605.13105v1 Announce Type: new Abstract: Reinforcement learning (RL) fine-tuning has shown promise for Vision-Language-Action (VLA) models in robotic manipulation, but deployment-time visual sh

safetyarxiv-cs-ro
14 May 2026
Safety

When Backdoors Meet Partial Observability: Attacking Real-World Reinforcement Learning

DGX agent

arXiv:2601.14104v2 Announce Type: replace-cross Abstract: Backdoor attacks can cause reinforcement learning (RL) policies to behave normally under clean inputs while executing malicious behaviors when

safetyarxiv-cs-cv
14 May 2026
Safety

When to Act, Ask, or Learn: Uncertainty-Aware Policy Steering

DGX agent

arXiv:2602.22474v2 Announce Type: replace-cross Abstract: Policy steering is an emerging way to adapt robot behaviors at deployment-time: a learned verifier analyzes low-level action samples proposed

safetyarxiv-cs-lg
14 May 2026
Safety

When to Trust Confidence Thresholding: Calibration Diagnostics for Pseudo-Labelled Regression

DGX agent

arXiv:2605.12780v1 Announce Type: cross Abstract: Calibrated probability outputs of trained classifiers are increasingly used as inputs to downstream regression estimands such as effects, prevalences,

safetyarxiv-cs-lg
14 May 2026
Safety

A Survey of On-Policy Distillation for Large Language Models

DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

safetyarxiv-cs-cl
13 May 2026
Safety

A Theory of Time-Sensitive Language Generation: Sparse Hallucination Beats Mode Collapse

DGX agent

arXiv:2605.11302v1 Announce Type: cross Abstract: We study language generation in the limit under a global preference ordering on strings, as introduced by Kleinberg and Wei. As in [arXiv:2504.14370,

safetyarxiv-cs-cl
13 May 2026
Safety

A Unified Graph Language Model for Multi-Domain Multi-Task Graph Alignment Instruction Tuning

DGX agent

arXiv:2605.12197v1 Announce Type: new Abstract: Leveraging Graph Neural Networks (GNNs) as graph encoders and aligning the resulting representations with Large Language Models (LLMs) through alignment

safetyarxiv-cs-lg
13 May 2026
Safety

ACSAC: Adaptive Chunk Size Actor-Critic with Causal Transformer Q-Network

DGX agent

arXiv:2605.11009v1 Announce Type: new Abstract: Long-horizon, sparse-reward tasks pose a fundamental challenge for reinforcement learning, since single-step TD learning suffers from bootstrapping erro

safetyarxiv-cs-lg
13 May 2026
Safety

Adaptive Policy Learning Under Unknown Network Interference

DGX agent

arXiv:2605.11191v1 Announce Type: cross Abstract: Adaptive experimentation under unknown network interference requires solving two coupled problems: (i) learning the underlying dynamics of interferenc

safetyarxiv-cs-lg
13 May 2026
Safety

Adaptive TD-Lambda for Cooperative Multi-agent Reinforcement Learning

DGX agent

arXiv:2605.11880v1 Announce Type: new Abstract: TD(lambda) in value-based MARL algorithms or the Temporal Difference critic learning in Actor-Critic-based (AC-based) algorithms synergistically integra

safetyarxiv-cs-lg
13 May 2026
Safety

Adaptive Teacher Exposure for Self-Distillation in LLM Reasoning

DGX agent

arXiv:2605.11458v1 Announce Type: cross Abstract: On-policy self-distillation has become a strong recipe for LLM reasoning, where a privileged teacher supervises the student's own rollouts while condi

safetyarxiv-cs-cl
13 May 2026
Safety

AIA: Rethinking Architecture Decoupling Strategy In Unified Multimodal Model

DGX agent

arXiv:2511.22663v5 Announce Type: replace Abstract: Unified multimodal models for image generation and understanding represent a significant step toward AGI and have attracted widespread attention fro

safetyarxiv-cs-cv
13 May 2026
Safety

Aligning Flow Map Policies with Optimal Q-Guidance

DGX agent

arXiv:2605.12416v1 Announce Type: new Abstract: Generative policies based on expressive model classes, such as diffusion and flow matching, are well-suited to complex control problems with highly mult

safetyarxiv-cs-lg
13 May 2026
Safety

AlphaGRPO: Unlocking Self-Reflective Multimodal Generation in UMMs via Decompositional Verifiable Reward

DGX agent

arXiv:2605.12495v1 Announce Type: new Abstract: In this paper, we propose AlphaGRPO, a novel framework that applies Group Relative Policy Optimization (GRPO) to AR-Diffusion Unified Multimodal Models

safetyarxiv-cs-cv
13 May 2026
Safety

Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information

DGX agent

arXiv:2605.11609v1 Announce Type: cross Abstract: On-policy self-distillation, where a student is pulled toward a copy of itself conditioned on privileged context (e.g., a verified solution or feedbac

safetyarxiv-cs-cl
13 May 2026
Safety

Assessment of cloud and associated radiation fields from a GAN stochastic cloud subcolumn generator

DGX agent

arXiv:2605.11968v1 Announce Type: cross Abstract: Modern Earth System Models (ESMs) operate on horizontal scales far larger than typical cloud features, requiring stochastic subcolumn generators to re

safetyarxiv-cs-lg
13 May 2026
Safety

Asymmetric Advantage Modulation Calibrates Entropy Dynamics in RLVR

DGX agent

arXiv:2604.04894v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has substantially improved the reasoning ability of large language models (LLMs), but it often

safetyarxiv-cs-cl
13 May 2026
Safety

Augmented Lagrangian Method for Last-Iterate Convergence for Constrained MDPs

DGX agent

arXiv:2605.11694v1 Announce Type: new Abstract: We study policy optimization for infinite-horizon, discounted constrained Markov decision processes (CMDPs). While existing theoretical guarantees typic

safetyarxiv-cs-lg
13 May 2026
Safety

Autoregressive Learning in Joint KL: Sharp Oracle Bounds and Lower Bounds

DGX agent

arXiv:2605.12316v1 Announce Type: new Abstract: We study the fundamental and timely problem of learning long sequences in autoregressive modeling and next-token prediction under model misspecification

safetyarxiv-cs-lg
13 May 2026
Safety

Beyond Point-wise Neural Collapse: A Topology-Aware Hierarchical Classifier for Class-Incremental Learning

DGX agent

arXiv:2605.11904v1 Announce Type: new Abstract: The Nearest Class Mean (NCM) classifier is widely favored in Class-Incremental Learning (CIL) for its superior resistance to catastrophic forgetting com

safetyarxiv-cs-cv
13 May 2026
Safety

Can Graphs Help Vision SSMs See Better?

DGX agent

arXiv:2605.11300v1 Announce Type: new Abstract: Vision state space models inherit the efficiency and long-range modeling ability of Mamba-style selective scans. However, their performance depends crit

safetyarxiv-cs-cv
13 May 2026
Safety

Causal Bias Detection in Generative Artifical Intelligence

DGX agent

arXiv:2605.11365v1 Announce Type: cross Abstract: Automated systems built on artificial intelligence (AI) are increasingly deployed across high-stakes domains, raising critical concerns about fairness

safetyarxiv-cs-lg
13 May 2026
Safety

Causal Fairness for Survival Analysis

DGX agent

arXiv:2605.11362v1 Announce Type: new Abstract: In the data-driven era, large-scale datasets are routinely collected and analyzed using machine learning (ML) and artificial intelligence (AI) to inform

safetyarxiv-cs-lg
13 May 2026
Safety

CheXTemporal: A Dataset for Temporally-Grounded Reasoning in Chest Radiography

DGX agent

arXiv:2605.11304v1 Announce Type: new Abstract: Chest radiograph interpretation requires temporal reasoning over prior and current studies, yet most vision-language models are trained on static image-

safetyarxiv-cs-cv
13 May 2026
Safety

Clarity: The Flexibility-Interpretability Trade-Off in Sparsity-aware Concept Bottleneck Models

DGX agent

arXiv:2601.21944v2 Announce Type: replace Abstract: The widespread adoption of deep learning models in computer vision has intensified concerns about interpretability. Despite strong performance, thes

safetyarxiv-cs-lg
13 May 2026
Safety

Cluster-Aware Neural Collapse Prompt Tuning for Long-Tailed Generalization of Vision-Language Models

DGX agent

arXiv:2605.11939v1 Announce Type: new Abstract: Prompt learning has emerged as an efficient alternative to fine-tuning pre-trained vision-language models (VLMs). Despite its promise, current methods s

safetyarxiv-cs-cv
13 May 2026
Safety

Combining On-Policy Optimization and Distillation for Long-Context Reasoning in Large Language Models

DGX agent

arXiv:2605.12227v1 Announce Type: new Abstract: Adapting large language models (LLMs) to long-context tasks requires post-training methods that remain accurate and coherent over thousands of tokens. E

safetyarxiv-cs-cl
13 May 2026
Safety

Controllable User Simulation

DGX agent

arXiv:2605.11519v1 Announce Type: cross Abstract: Using offline datasets to evaluate conversational agents often fails to cover rare scenarios or to support testing new policies. This has motivated th

safetyarxiv-cs-cl
13 May 2026
Safety

Coordinated Diffusion: Generating Multi-Agent Behavior Without Multi-Agent Demonstrations

DGX agent

arXiv:2605.11485v1 Announce Type: new Abstract: Imitation learning powered by generative models has proven effective for modeling complex single-agent behaviors. However, teaching multi-agent systems,

safetyarxiv-cs-ro
13 May 2026
← Previous
1…184185186187188…260
Next →