AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Miner:Mining Intrinsic Mastery for Data-Efficient RL in Large Reasoning Models

DGX agent

arXiv:2601.04731v2 Announce Type: replace Abstract: Current critic-free RL methods for large reasoning models suffer from severe inefficiency when training on positive homogeneous prompts (where all r

safetyarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

MoCoTalk: Multi-Conditional Diffusion with Adaptive Router for Controllable Talking Head Generation

DGX agent

arXiv:2605.08050v1 Announce Type: new Abstract: Talking-head generation requires joint modeling of identity, head pose, facial expression, and mouth dynamics. Existing methods typically address only a

safetyarxiv-cs-cv
11 May 2026
Safety

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models

DGX agent

arXiv:2602.07026v2 Announce Type: replace-cross Abstract: Despite the success of multimodal contrastive learning in aligning visual and linguistic representations, a persistent geometric anomaly, the

safetyarxiv-cs-ai
11 May 2026
Safety

MPD^2-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis

DGX agent

arXiv:2605.08024v1 Announce Type: new Abstract: Learning-to-defer (L2D) can make glaucoma screening safer by routing difficult/uncertain cases to humans, yet standard formulations overlook expert avai

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-environment Invariance Learning with Missing Data

DGX agent

arXiv:2601.07247v2 Announce Type: replace-cross Abstract: Learning models that can handle distribution shifts is a key challenge in domain generalization. Invariance learning, an approach that focuses

safetyarxiv-cs-lg
11 May 2026
Safety

Multi-Environment POMDPs with Finite-Horizon Objectives

DGX agent

arXiv:2605.07537v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are systems in which one agent interacts with a stochastic environment, and receives only partia

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation

DGX agent

arXiv:2603.16876v2 Announce Type: replace-cross Abstract: We propose MARL-Rad, a multi-modal multi-agent reinforcement learning framework for radiology report generation that trains the entire agentic

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Objective Multi-Agent Bandits: From Learning Efficiency to Fairness Optimization

DGX agent

arXiv:2605.06864v1 Announce Type: new Abstract: We study multi-objective multi-agent multi-armed bandits (MO-MA-MAB) under stochastic rewards, where agents observe heterogeneous reward vectors and com

safetyarxiv-cs-lg
11 May 2026
Safety

No Forgetting Learning: Buffer-free Continual Learning Classification

DGX agent

arXiv:2503.04638v3 Announce Type: replace Abstract: Most Continual Learning (CL) methods maintain performance on earlier tasks by storing exemplars in a replay buffer, introducing memory overhead that

safetyarxiv-cs-lg
11 May 2026
Safety

NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models

DGX agent

arXiv:2605.07794v1 Announce Type: new Abstract: World Action Models (WAMs) are an emerging family of policies that tie robot action generation to future-observation modeling. In this work, we focus on

safetyarxiv-cs-ro
11 May 2026
Safety

OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search

DGX agent

arXiv:2604.03675v2 Announce Type: replace Abstract: Agentic search enables language models to solve knowledge-intensive tasks by adaptively acquiring external evidence over multiple steps. Reinforceme

safetyarxiv-cs-ai
11 May 2026
Safety

Object Hallucination-Free Reinforcement Unlearning for Vision-Language Models

DGX agent

arXiv:2605.08031v1 Announce Type: new Abstract: Vision-language models (VLMs) raise growing concerns about privacy, copyright, and bias, motivating machine unlearning to remove sensitive knowledge. Ho

safetyarxiv-cs-cv
11 May 2026
Safety

Offline Policy Optimization with Posterior Sampling

DGX agent

arXiv:2605.07393v1 Announce Type: new Abstract: A fundamental challenge in model-based offline reinforcement learning (RL) lies in the trade-off between generalization and robustness against exploitat

safetyarxiv-cs-ai
11 May 2026
Safety

On the Meta-Design of Allocation Problems

DGX agent

arXiv:2602.08786v4 Announce Type: replace-cross Abstract: There is an extensive literature that studies how to find optimal policies in resource allocation problems, taking the underlying design param

safetyarxiv-cs-lg
11 May 2026
Safety

On Training in Imagination

DGX agent

arXiv:2605.06732v1 Announce Type: new Abstract: State-of-the-art model-based reinforcement learning methods train policies on imagined rollouts. These rollouts are trajectories generated by a learned

safetyarxiv-cs-lg
11 May 2026
Safety

One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy

DGX agent

arXiv:2605.07931v1 Announce Type: cross Abstract: Vision-language-action (VLA) models increasingly rely on auxiliary world modules to plan over long horizons, yet how such modules should be parameteri

safetyarxiv-cs-ai
11 May 2026
Safety

Online Allocation with Unknown Shared Supply

DGX agent

arXiv:2605.07080v1 Announce Type: new Abstract: Many real-world resource allocation systems, such as humanitarian logistics and vaccine distribution, must preposition limited supply across multiple lo

safetyarxiv-cs-ai
11 May 2026
Safety

Optimal Recourse Summaries via Bi-Objective Decision Tree Learning

DGX agent

arXiv:2605.07598v1 Announce Type: new Abstract: Actionable Recourse provides individuals with actions they can take to change an unfavorable classifier outcome. While useful at the instance level, it

safetyarxiv-cs-lg
11 May 2026
Safety

PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents

DGX agent

arXiv:2605.07039v1 Announce Type: new Abstract: Large language models have become drivers of evolutionary search, but most systems rely on a fixed, prompt-elicited policy to sample next candidates. Th

safetyarxiv-cs-lg
11 May 2026
Safety

Pan-FM: A Pan-Organ Foundation Model with Saliency-Guided Masking for Missing Robustness

DGX agent

arXiv:2605.07055v1 Announce Type: cross Abstract: Foundation models (FMs) have shown great promise in medical imaging, but most FMs are trained on unimodal data within isolated domains, such as brain

safetyarxiv-cs-ai
11 May 2026
Safety

PaT: Planning-after-Trial for Efficient Test-Time Code Generation

DGX agent

arXiv:2605.07248v1 Announce Type: new Abstract: Beyond training-time optimization, scaling test-time computation has emerged as a key paradigm to extend the reasoning capabilities of Large Language Mo

safetyarxiv-cs-cl
11 May 2026
Safety

Persistent-Transient Policy Evaluation for Markov Chains via Minimal Peripheral Quotients

DGX agent

arXiv:2602.00474v2 Announce Type: replace-cross Abstract: We study fixed-policy evaluation for finite Markov chains that may be reducible and periodic. Classical evaluation methods with gain and bias

safetyarxiv-cs-lg
11 May 2026
Safety

Physical Simulators as Do-Operators: Causal Discovery under Latent Confounders for AI-for-Science

DGX agent

arXiv:2605.07467v1 Announce Type: cross Abstract: Existing interventional causal discovery methods -- IGSP, DCDI, ENCO -- assume causal sufficiency (no latent confounders) and rely on virtual interven

safetyarxiv-cs-ai
11 May 2026
Safety

Physics-Based Benchmarking Metrics for Multimodal Synthetic Images

DGX agent

arXiv:2511.15204v3 Announce Type: replace-cross Abstract: Current state of the art measures like BLEU, CIDEr, VQA score, SigLIP-2 and CLIPScore are often unable to capture semantic or structural accur

safetyarxiv-cs-ai
11 May 2026
Safety

PLOT: Progressive Localization via Optimal Transport in Neural Causal Abstraction

DGX agent

arXiv:2605.06979v1 Announce Type: cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning a high-level causal model with the low-level computation r

safetyarxiv-cs-ai
11 May 2026
Safety

POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles

DGX agent

arXiv:2605.07775v1 Announce Type: cross Abstract: Balancing exploration and exploitation is a core challenge in sequential decision-making and black-box optimization. We introduce POETS (extbf{Po}licy

safetyarxiv-cs-ai
11 May 2026
Safety

Position: Mechanistic Interpretability Must Disclose Identification Assumptions for Causal Claims

DGX agent

arXiv:2605.08012v1 Announce Type: cross Abstract: Mechanistic interpretability papers increasingly use causal vocabulary: circuits, mediators, causal abstraction, monosemanticity. Such claims require

safetyarxiv-cs-ai
11 May 2026
Safety

Post-training makes large language models less human-like

DGX agent

arXiv:2605.07632v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavi

safetyarxiv-cs-ai
11 May 2026
Safety

ProtoSSL: Interpretable Prototype Learning from Unlabeled Time-Series Data

DGX agent

arXiv:2605.06943v1 Announce Type: new Abstract: In time-series domains where both predictive performance and interpretability are essential, deep neural networks achieve strong results but provide lim

safetyarxiv-cs-lg
11 May 2026
Safety

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment

DGX agent

arXiv:2605.08064v1 Announce Type: new Abstract: Spatial intelligence in vision-language models (VLMs) attracts research interest with the practical demand to reason in the 3D world.Despite promising r

safetyarxiv-cs-cv
11 May 2026
Safety

Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning

DGX agent

arXiv:2605.07804v1 Announce Type: cross Abstract: On-policy distillation (OPD) leverages dense teacher rewards to enhance reasoning models. However, scaling OPD to long-horizon tasks exposes a critica

safetyarxiv-cs-ai
11 May 2026
Safety

Q-MMR: Off-Policy Evaluation via Recursive Reweighting and Moment Matching

DGX agent

arXiv:2605.06474v2 Announce Type: replace-cross Abstract: We present a novel theoretical framework, Q-MMR, for off-policy evaluation in finite-horizon MDPs. Q-MMR learns a set of scalar weights, one f

safetyarxiv-cs-ai
11 May 2026
Safety

R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes

DGX agent

arXiv:2601.20599v2 Announce Type: replace-cross Abstract: Gradient temporal-difference (GTD) learning algorithms are widely used for off-policy policy evaluation with function approximation. However,

safetyarxiv-cs-ai
11 May 2026
Safety

Radiologist-Guided Causal Concept Bottleneck Models for Chest X-Ray Interpretation

DGX agent

arXiv:2605.07785v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) in medical imaging aim to improve model interpretability by predicting intermediate clinical concepts before final diag

safetyarxiv-cs-cv
11 May 2026
Safety

Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners

DGX agent

arXiv:2605.08019v1 Announce Type: new Abstract: Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent actio

safetyarxiv-cs-ai
11 May 2026
Safety

ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning

DGX agent

arXiv:2605.07477v1 Announce Type: new Abstract: Recent text-guided image editing (TIE) models have achieved remarkable progress, however, many edited results still suffer from artifacts, unintended mo

safetyarxiv-cs-cv
11 May 2026
Safety

ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation

DGX agent

arXiv:2408.06747v4 Announce Type: replace Abstract: Recent works utilize CLIP to perform the challenging unsupervised semantic segmentation task where only images without annotations are available. Ho

safetyarxiv-cs-cv
11 May 2026
Safety

Reflections and New Directions for Human-Centered Large Language Models

DGX agent

arXiv:2605.06901v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping the private and professional lives of users, with numerous applications in business, education, fi

safetyarxiv-cs-cl
11 May 2026
Safety

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs

DGX agent

arXiv:2605.08053v1 Announce Type: new Abstract: Reinforcement learning (RL) for exponential-utility optimization in discounted Markov decision processes (MDPs) lacks principled value-based algorithms.

safetyarxiv-cs-lg
11 May 2026
Safety

RELO: Reinforcement Learning to Localize for Visual Object Tracking

DGX agent

arXiv:2605.07379v1 Announce Type: cross Abstract: Conventional visual object trackers localize targets using handcrafted spatial priors, often in the form of heatmaps. Such priors provide only surroga

safetyarxiv-cs-ai
11 May 2026
Safety

Repeated Deceptive Path Planning against Learnable Observer

DGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

safetyarxiv-cs-ai
11 May 2026
Safety

Resource-Element Energy Difference for Noncoherent Over-the-Air Federated Learning

DGX agent

arXiv:2605.07263v1 Announce Type: cross Abstract: Over-the-air federated learning (OTA-FL) reduces uplink latency by exploiting waveform superposition, but conventional analog aggregation schemes typi

safetyarxiv-cs-ai
11 May 2026
Safety

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding

DGX agent

arXiv:2605.07575v1 Announce Type: cross Abstract: Proactive streaming video understanding requires Video-LLMs to decide when to respond as a video unfolds, a task where existing methods often fall sho

safetyarxiv-cs-ai
11 May 2026
Safety

Response Time Enhances Alignment with Heterogeneous Preferences

DGX agent

arXiv:2605.06987v1 Announce Type: new Abstract: Aligning large language models (LLMs) to human preferences typically relies on aggregating pooled feedback into a single reward model. However, this sta

safetyarxiv-cs-lg
11 May 2026
Safety

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

DGX agent

arXiv:2605.07331v1 Announce Type: cross Abstract: Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Cen

safetyarxiv-cs-ai
11 May 2026
Safety

RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion

DGX agent

arXiv:2602.16548v2 Announce Type: replace Abstract: The inverse design of RNA three-dimensional (3D) structures is crucial for engineering functional RNAs in synthetic biology and therapeutics. While

safetyarxiv-cs-lg
11 May 2026
Safety

Risk-Consistent Multiclass Learning from Random Label-Subset Membership Queries

DGX agent

arXiv:2605.07413v1 Announce Type: new Abstract: Obtaining accurate class labels is often costly or unreliable, and may also be limited by privacy or other practical conditions. Compared with asking an

safetyarxiv-cs-lg
11 May 2026
Safety

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

DGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

safetyarxiv-cs-lg
11 May 2026
← Previous
1…194195196197198…257
Next →