AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Rollback-Free Stable Brick Structures Generation

DGX agent

arXiv:2605.06947v1 Announce Type: new Abstract: While autoregressive models have advanced 3D generation, creating physically stable brick structures remains a challenge due to the strict requirements

safetyarxiv-cs-lg
11 May 2026
Safety

Rubric-based On-policy Distillation

DGX agent

arXiv:2605.07396v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a powerful paradigm for model alignment, yet its reliance on teacher logits restricts its application to white-box sce

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
safetyarxiv-cs-ai
11 May 2026
Safety

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

DGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

safetyarxiv-cs-ai
11 May 2026
Safety

SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions

DGX agent

arXiv:2605.07102v1 Announce Type: new Abstract: Evaluating literary quality requires assessing interpretive dimensions such as cultural representation, emotional depth, and philosophical sophisticatio

safetyarxiv-cs-cl
11 May 2026
Safety

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

DGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

safetyarxiv-cs-ai
11 May 2026
Safety

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

DGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

safetyarxiv-cs-cv
11 May 2026
Safety

Self-Programmed Execution for Language-Model Agents

DGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

safetyarxiv-cs-ai
11 May 2026
Safety

SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents

DGX agent

arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa

safetyarxiv-cs-lg
11 May 2026
Safety

SimCT: Recovering Lost Supervision for Cross-Tokenizer On-Policy Distillation

DGX agent

arXiv:2605.07711v1 Announce Type: new Abstract: On-policy distillation (OPD) is a standard tool for transferring teacher behavior to a smaller student, but it implicitly assumes that teacher and stude

safetyarxiv-cs-cl
11 May 2026
Safety

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

DGX agent

arXiv:2605.06130v2 Announce Type: replace Abstract: A persistent skill library allows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupl

safetyarxiv-cs-ai
11 May 2026
Safety

Slowly Annealed Langevin Dynamics: Theory and Applications to Training-Free Guided Generation

DGX agent

arXiv:2605.07950v1 Announce Type: new Abstract: We study Slowly Annealed Langevin Dynamics (SALD), a sampler for tracking a path of moving target distributions and approximating the terminal target th

safetyarxiv-cs-lg
11 May 2026
Safety

SOD: Step-wise On-policy Distillation for Small Language Model Agents

DGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

safetyarxiv-cs-ai
11 May 2026
Safety

SparseRL-Sync: Lossless Weight Synchronization with ~100x Less Communication

DGX agent

arXiv:2605.07330v1 Announce Type: cross Abstract: In large-scale reinforcement learning (RL) systems with decoupled Trainer-Rollout execution, the Trainer must regularly synchronize policy weights to

safetyarxiv-cs-ai
11 May 2026
Safety

Stabilized neural Hamilton--Jacobi--Bellman solvers: Error analysis and applications in model-based reinforcement learning

DGX agent

arXiv:2605.07116v1 Announce Type: cross Abstract: Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is

safetyarxiv-cs-ai
11 May 2026
Safety

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration

DGX agent

arXiv:2512.23927v2 Announce Type: replace-cross Abstract: Fitted Q-iteration (FQI) and soft FQI are widely used value-based methods for offline reinforcement learning, but their standard stability gua

safetyarxiv-cs-lg
11 May 2026
Safety

STDA-Net: Spectrogram-Based Domain Adaptation for cross-dataset Sleep Stage Classification

DGX agent

arXiv:2605.06736v1 Announce Type: cross Abstract: Accurate sleep stage classification across datasets remains challenging due to variability in EEG channel montages, sampling rates, recording environm

safetyarxiv-cs-ai
11 May 2026
Safety

Structured Role-Aware Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2605.07274v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR), especially with Group Relative Policy Optimization (GRPO), has shown strong potential for improvi

safetyarxiv-cs-ai
11 May 2026
Safety

Supervised sparse auto-encoders for interpretable and compositional representations

DGX agent

arXiv:2602.00924v2 Announce Type: replace Abstract: Sparse auto-encoders (SAEs) have re-emerged as a prominent method for mechanistic interpretability, yet they face two significant challenges: the no

safetyarxiv-cs-ai
11 May 2026
Safety

Temporal Attention for Adaptive Control of Euler-Lagrange Systems with Unobservable Memory

DGX agent

arXiv:2605.06877v1 Announce Type: new Abstract: Adaptive control of Euler-Lagrange systems is challenging when friction is governed by a finite-horizon internal state that is not directly observable f

safetyarxiv-cs-lg
11 May 2026
Safety

Temporal Smoothness Doubly Robust Learning for Debiased Knowledge Tracing

DGX agent

arXiv:2605.05958v2 Announce Type: replace Abstract: Knowledge Tracing (KT) is fundamental to intelligent education systems, yet relies on educational logs that are selectively observed. The non-random

safetyarxiv-cs-ai
11 May 2026
Safety

TextLDM: Language Modeling with Continuous Latent Diffusion

DGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

safetyarxiv-cs-cl
11 May 2026
Safety

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

DGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

safetyarxiv-cs-ro
11 May 2026
Safety

The Effect of Mini-Batch Noise on the Implicit Bias of Adam

DGX agent

arXiv:2602.01642v2 Announce Type: replace-cross Abstract: With limited high-quality data and growing compute, multi-epoch training is gaining back its importance across sub-areas of deep learning. Ada

safetyarxiv-cs-ai
11 May 2026
Safety

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

DGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

safetyarxiv-cs-ai
11 May 2026
Safety

The Limits of AI-Driven Allocation: Optimal Screening under Aleatoric Uncertainty

DGX agent

arXiv:2605.07979v1 Announce Type: new Abstract: The rise of machine learning has shifted targeted resource allocation in policy and humanitarian settings toward algorithmic targeting based on predicte

safetyarxiv-cs-ai
11 May 2026
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Safety

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

DGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

safetyarxiv-cs-cl
11 May 2026
Safety

Toward Better Geometric Representations for Molecule Generative Models

DGX agent

arXiv:2605.07693v1 Announce Type: new Abstract: Geometric representation-conditioned molecule generation provides an effective paradigm that decouples molecule representation modeling from structure g

safetyarxiv-cs-lg
11 May 2026
Safety

Towards Differentially Private Reinforcement Learning with General Function Approximation

DGX agent

arXiv:2605.07049v1 Announce Type: cross Abstract: We present the first theoretical guarantees for differentially private online reinforcement learning (RL) with general function approximation, extendi

safetyarxiv-cs-ai
11 May 2026
Safety

Towards Fairness under Label Bias in Image Segmentation: Impact, Measurement and Mitigation

DGX agent

arXiv:2605.06891v1 Announce Type: new Abstract: Labeled datasets reflect the biases of their annotation pipelines, which sometimes introduce label bias: group-conditional label errors that cause syste

safetyarxiv-cs-cv
11 May 2026
Safety

TRACE: Transport Alignment Conformal Prediction via Diffusion and Flow Matching Models

DGX agent

arXiv:2605.07100v1 Announce Type: cross Abstract: Constructing valid and informative conformal prediction regions for multi-dimensional outputs remains a fundamental challenge. While conformal predict

safetyarxiv-cs-lg
11 May 2026
Safety

Training-Free Multimodal Large Language Model Orchestration

DGX agent

arXiv:2508.10016v3 Announce Type: replace Abstract: Building interactive omni-modal assistants often relies on end-to-end multimodal alignment to fuse heterogeneous modalities, which incurs substantia

safetyarxiv-cs-cl
11 May 2026
Safety

TRAJGANR: Trajectory-Centric Urban Multimodal Learning via Geospatially Aligned Neural Representations

DGX agent

arXiv:2605.06990v1 Announce Type: new Abstract: Multimodal self-supervised learning (MSSL) has emerged as a key paradigm for pretraining geospatial foundation models. However, existing geospatial MSSL

safetyarxiv-cs-cv
11 May 2026
Safety

UFT: Unifying Fine-Tuning of SFT and RLHF/DPO/UNA through a Generalized Implicit Reward Function

DGX agent

arXiv:2410.21438v3 Announce Type: replace Abstract: By pretraining on trillions of tokens, an LLM gains the capability of text generation. However, to enhance its utility and reduce potential harm, SF

safetyarxiv-cs-cl
11 May 2026
Safety

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types

DGX agent

arXiv:2408.15339v4 Announce Type: replace-cross Abstract: RL alignment methods, including RLHF and DPO, are primarily based on pairwise preference data. Although scalar or score-based feedback has bee

safetyarxiv-cs-cl
11 May 2026
Safety

UniD-Shift: Towards Unified Semantic Segmentation via Interpretable Share-Private Multimodal Decomposition

DGX agent

arXiv:2605.07356v1 Announce Type: new Abstract: Semantic segmentation of large-scale 3D point clouds is crucial for applications such as autonomous driving and urban digital twins. However, the sparse

safetyarxiv-cs-cv
11 May 2026
Safety

VDEGaussian: Video Diffusion Enhanced 4D Gaussian Splatting for Dynamic Urban Scenes Modeling

DGX agent

arXiv:2508.02129v2 Announce Type: replace Abstract: Dynamic urban scene modeling is a rapidly evolving area with broad applications. While current approaches leveraging neural radiance fields or Gauss

safetyarxiv-cs-cv
11 May 2026
Safety

VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training

DGX agent

arXiv:2602.10693v3 Announce Type: replace-cross Abstract: Off-policy updates are inevitable in reinforcement learning (RL) for large language models (LLMs) due to rollout staleness from asynchronous t

safetyarxiv-cs-ai
11 May 2026
Safety

VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding

DGX agent

arXiv:2605.05848v2 Announce Type: replace-cross Abstract: Video large multimodal models increasingly face a scalability bottleneck: long videos produce excessively long visual-token sequences, which s

safetyarxiv-cs-ai
11 May 2026
Safety

VISD: Enhancing Video Reasoning via Structured Self-Distillation

DGX agent

arXiv:2605.06094v2 Announce Type: replace-cross Abstract: Training VideoLLMs for complex reasoning remains challenging due to sparse sequence level rewards and the lack of fine grained credit assignme

safetyarxiv-cs-ai
11 May 2026
Safety

When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize

DGX agent

arXiv:2605.06868v1 Announce Type: new Abstract: Fixed-budget nonconvex optimization can fail not because local descent is unstable, but because it is too stable: after reaching a nearby stationary poi

safetyarxiv-cs-lg
11 May 2026
Safety

Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

DGX agent

arXiv:2605.05558v2 Announce Type: replace Abstract: A natural intuition about the economics of AI agents is that, because agents can be replicated at very low marginal cost, agent labor may be supplie

safetyarxiv-cs-ai
11 May 2026
Safety

Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping

DGX agent

arXiv:2605.08075v1 Announce Type: new Abstract: Decoding imagined speech from non-invasive brain recordings is challenging because imagined datasets are scarce and difficult to align temporally across

safetyarxiv-cs-lg
11 May 2026
Safety

A cross-modal network for facial expression recognition

DGX agent

arXiv:2605.04439v1 Announce Type: new Abstract: Deep neural networks enriched with structural information have been widely employed for facial expression recognition tasks. However, these methods ofte

safetyarxiv-cs-cv
7 May 2026
Safety

A Skill-Based AI Agentic Pipeline for Library of Congress Subject Indexing

DGX agent

arXiv:2605.03537v1 Announce Type: cross Abstract: This paper presents a modular AI agentic skill pipeline for automating subject indexing with Library of Congress Subject Headings (LCSH). Subject inde

safetyarxiv-cs-ai
7 May 2026
Safety

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning

DGX agent

arXiv:2605.04066v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an essential paradigm that enhances the reasoning capabilities of Large Language Models (LLMs).

safetyarxiv-cs-cl
7 May 2026
Safety

Adaptive Dual-Path Framework for Covert Semantic Communication

DGX agent

arXiv:2605.03423v1 Announce Type: new Abstract: This paper proposes a novel adaptive dual-path framework for covert semantic communication (SemCom), which integrates covert information transmission wi

safetyarxiv-cs-ai
7 May 2026
Safety

Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2605.05123v1 Announce Type: new Abstract: In offline-to-online reinforcement learning (O2O-RL), policies are first safely trained offline using previously collected datasets and then further fin

safetyarxiv-cs-lg
7 May 2026
← Previous
1…195196197198199…257
Next →