AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Precise: SDE-Consistent Stochastic Sampling for RL Post-Training of Flow-Matching Models

DGX agent

arXiv:2605.23522v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become an effective way to improve prompt alignment and perceptual quality in diffusion and flow-matching generators.

safetyarxiv-cs-ai
25 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Pure Exploration for a Good Policy in Reinforcement Learning with Bandit Feedback

DGX agent

arXiv:2605.23182v1 Announce Type: new Abstract: Pure exploration in episodic Reinforcement Learning has primarily focused on Best Policy Identification (BPI), which seeks to identify a (near)-optimal

safetyarxiv-cs-lg
25 May 2026
Safety

RAG4Outcome: A Retrieval-Augmented Multimodal Framework for Prognostic Prediction in Chronic Osteomyelitis

DGX agent

arXiv:2605.22833v1 Announce Type: cross Abstract: Chronic osteomyelitis presents substantial prognostic challenges due to its high recurrence risk and complex postoperative recovery trajectories. Trad

safetyarxiv-cs-ai
25 May 2026
Safety

ReCoVer: Resilient LLM Pre-Training System via Fault-Tolerant Collective and Versatile Workload

DGX agent

arXiv:2605.11215v2 Announce Type: replace-cross Abstract: Pre-training large language models on massive GPU clusters has made hardware faults routine rather than rare, driving the need for resilient t

safetyarxiv-cs-ai
25 May 2026
Safety

Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control

DGX agent

arXiv:2605.23415v1 Announce Type: cross Abstract: Reinforcement learning has long struggled with poor sample efficiency. One promising approach to mitigate this problem is leveraging group-invariant M

safetyarxiv-cs-ai
25 May 2026
Safety

Reinforcement Learning with Verifiable yet Noisy Rewards under Imperfect Verifiers

DGX agent

arXiv:2510.00915v4 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) replaces costly human labeling with automated verifiers. To reduce verifier hacking, man

safetyarxiv-cs-ai
25 May 2026
Safety

Robotic Strawberry Harvesting with Robust Vision and Deep Reinforcement Learning based Sim-to-Real Control

DGX agent

arXiv:2605.23863v1 Announce Type: new Abstract: This study presents a closed-loop robotic strawberry harvesting system that combines a robust vision module, simulation-trained deep reinforcement learn

safetyarxiv-cs-ro
25 May 2026
Safety

Sample-wise Targeted Adversarial Attacks on Test-time Adaptation

DGX agent

arXiv:2605.23411v1 Announce Type: cross Abstract: Test-time adaptation (TTA) effectively counters distribution shifts but exposes models to adversarial manipulation via the unlabeled test stream. Exis

safetyarxiv-cs-cv
25 May 2026
Safety

Score-Based One-step MeanFlow Policy Optimization

DGX agent

arXiv:2605.23365v1 Announce Type: cross Abstract: Diffusion and flow matching have emerged as expressive policy classes in reinforcement learning, but their reliance on multi-step denoising imposes su

safetyarxiv-cs-ai
25 May 2026
Safety

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-Based Humanoid Control

DGX agent

arXiv:2605.22894v1 Announce Type: cross Abstract: Controlling physics-based humanoids from natural-language instructions is a critical step toward general-purpose embodied agents. However, existing me

safetyarxiv-cs-lg
25 May 2026
Safety

SeedER: Seed-and-Expand Retrieval from Knowledge Graphs

DGX agent

arXiv:2605.23753v1 Announce Type: new Abstract: Knowledge graphs (KGs) offer a rich representation for relational knowledge, but their irregular structure makes retrieval challenging: ego-graph expans

safetyarxiv-cs-lg
25 May 2026
Safety

Spatio-Temporal Similarity Volume Aggregation for Open-Vocabulary Action Recognition

DGX agent

arXiv:2605.23288v1 Announce Type: new Abstract: Recent Open-Vocabulary Action Recognition (OVAR) methods typically aggregate visual features into a global representation before computing text alignmen

safetyarxiv-cs-cv
25 May 2026
Safety

SpinFlow: A Physics-Informed Spin Field Framework for Traffic Phase Inference and Transition Detection

DGX agent

arXiv:2605.23306v1 Announce Type: cross Abstract: Active traffic management (ATM) is frequently hindered by traditional macroscopic models and rigid empirical thresholds that fail to capture metastabl

safetyarxiv-cs-lg
25 May 2026
Safety

Task-Awareness Improves LLM Generations and Uncertainty

DGX agent

arXiv:2601.21500v2 Announce Type: replace Abstract: In many applications of LLMs, natural language responses often have an underlying structure such as representing discrete labels, numerical values,

safetyarxiv-cs-lg
25 May 2026
Safety

The Implicit Bias of Depth: From Neural Collapse to Softmax Codes

DGX agent

arXiv:2605.23087v1 Announce Type: new Abstract: Neural collapse (NC) describes the structured geometry that emerges in the features and weights of trained classifiers. Recent theory suggests NC can be

safetyarxiv-cs-lg
25 May 2026
Safety

The physics of AI weather models

DGX agent

arXiv:2605.23778v1 Announce Type: cross Abstract: Could it be that AI weather models are solving physical equations, although they may not be the equations used by conventional NWP models? We compute

safetyarxiv-cs-lg
25 May 2026
Safety

Through the Stealth Lens: Attention-Aware Defenses Against Poisoning in RAG

DGX agent

arXiv:2506.04390v2 Announce Type: replace-cross Abstract: Retrieval-augmented generation (RAG) systems are vulnerable to attacks that inject poisoned passages into the retrieved context, even at low c

safetyarxiv-cs-ai
25 May 2026
Safety

Transform-Invariant Generative Ray Path Sampling for Efficient Radio Propagation Modeling

DGX agent

arXiv:2603.01655v2 Announce Type: replace Abstract: Ray tracing has become a standard for accurate radio propagation modeling, but suffers from exponential computational complexity, as the number of c

safetyarxiv-cs-lg
25 May 2026
Safety

Understanding Goal Generalisation in Sequential Reinforcement Learning

DGX agent

arXiv:2605.23565v1 Announce Type: cross Abstract: Reinforcement learning agents often exhibit unintended goal-directed behaviour outside their training distribution, but we currently lack a principled

safetyarxiv-cs-ai
25 May 2026
Safety

UniEmo: Unifying Emotional Understanding and Generation with Learnable Expert Queries

DGX agent

arXiv:2507.23372v2 Announce Type: replace Abstract: Emotional understanding and generation are often treated as separate tasks, yet they are inherently complementary and can mutually enhance each othe

safetyarxiv-cs-cv
25 May 2026
Safety

UniReg: A Universal Model for Controllable CT Image Registration

DGX agent

arXiv:2503.12868v2 Announce Type: replace Abstract: Learning-based medical image registration has matched the accuracy of conventional methods while offering superior computational efficiency. However

safetyarxiv-cs-cv
25 May 2026
Safety

V-VLAPS: Value-Guided Planning for Vision-Language-Action Models

DGX agent

arXiv:2601.00969v2 Announce Type: replace-cross Abstract: Vision-language-action (VLA) models provide strong action priors for robotic manipulation, but their reactive behavior can fail under distribu

safetyarxiv-cs-ai
25 May 2026
Safety

VI-CuRL: Stabilizing Verifier-Independent RL Reasoning via Confidence-Guided Variance Reduction

DGX agent

arXiv:2602.12579v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a dominant paradigm for enhancing Large Language Models (LLMs) reasoning,

safetyarxiv-cs-ai
25 May 2026
Safety

Vision Transformers Need Better Token Interaction

DGX agent

arXiv:2605.23868v1 Announce Type: new Abstract: Vision Transformers (ViTs) can learn strong image-level representations while their patch representations become less effective for dense prediction dur

safetyarxiv-cs-cv
25 May 2026
Safety

When Determinants Are Not Enough: Private Rare Switching

DGX agent

arXiv:2605.23131v1 Announce Type: new Abstract: In this note, I would like to share a small research moment where Codex helped me find the right way to adapt rare switching to the private setting. The

safetyarxiv-cs-lg
25 May 2026
Safety

Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good

DGX agent

arXiv:2605.22995v1 Announce Type: cross Abstract: Agentic AI systems are increasingly proposed for social-good domains, often invoking the United Nations Sustainable Development Goals (SDGs) as a voca

safetyarxiv-cs-ai
25 May 2026
Safety

A note on convergence of Wasserstein policy optimization

DGX agent

arXiv:2605.22622v1 Announce Type: new Abstract: Wasserstein Policy Optimization (WPO) is a recently proposed reinforcement learning algorithm that leverages Wasserstein gradient flows to optimize stoc

safetyarxiv-cs-lg
23 May 2026
Safety

A Tale of Two Cities: Pessimism and Opportunism in Offline Dynamic Pricing

DGX agent

arXiv:2411.08126v2 Announce Type: replace-cross Abstract: We study offline dynamic pricing when historical data provide incomplete coverage of the price space such that some candidate prices, includin

safetyarxiv-cs-lg
23 May 2026
Safety

Algebraic Machine Learning for Small-to-Medium Datasets Is Competitive against Strong Standard Baselines

DGX agent

arXiv:2605.22155v1 Announce Type: new Abstract: Symbolic methods are generally not considered competitive with strong modern learners on realistic supervised tasks. We evaluate Algebraic Machine Learn

safetyarxiv-cs-lg
23 May 2026
Safety

ARC-STAR: Auditable Post-Hoc Correction for PDE Foundation Models

DGX agent

arXiv:2605.22222v1 Announce Type: new Abstract: Partial differential equation (PDE) foundation models are pretrained networks that forecast how physical fields like velocity and pressure evolve from a

safetyarxiv-cs-lg
23 May 2026
Safety

Beyond One-Size-Fits-All: Adaptive Subgraph Denoising for Zero-Shot Graph Learning with Large Language Models

DGX agent

arXiv:2603.02938v2 Announce Type: replace Abstract: Graph-based tasks in the zero-shot setting remain a significant challenge due to data scarcity and the inability of traditional Graph Neural Network

safetyarxiv-cs-lg
23 May 2026
Safety

BioFormer: Rethinking Cross-Subject Generalization via Spectral Structural Alignment in Biomedical Time-Series

DGX agent

arXiv:2605.22468v1 Announce Type: new Abstract: Cross-subject generalization in biomedical time-series refers to training on data from some subjects and testing on unseen subjects.The key challenge is

safetyarxiv-cs-lg
23 May 2026
Safety

Causal Discovery in Structural VAR Models Under Equal Noise Variance

DGX agent

arXiv:2605.21846v1 Announce Type: cross Abstract: Causal discovery from multivariate time series is challenging when causal effects may occur both across time and within the same sampling interval. Th

safetyarxiv-cs-lg
23 May 2026
Safety

CCLab: Adversarial Testing of Learning- and Non-Learning-Based Congestion Controllers

DGX agent

arXiv:2605.21915v1 Announce Type: cross Abstract: Congestion controllers (CCs) are critical to network performance, and yet their robustness under adverse conditions remains insufficiently understood.

safetyarxiv-cs-lg
23 May 2026
Safety

Corruption-Tolerant Asynchronous Q-Learning with Near-Optimal Rates

DGX agent

arXiv:2509.08933v2 Announce Type: replace Abstract: We study the problem of learning the optimal policy in a discounted, infinite-horizon reinforcement learning (RL) setting in the presence of adversa

safetyarxiv-cs-lg
23 May 2026
Safety

Cross-Species RSA Reveals Conserved Early Visual Alignment but Divergent Higher-Area Rankings Across Human fMRI and Macaque Electrophysiology

DGX agent

arXiv:2605.22401v1 Announce Type: new Abstract: Does the relationship between learning rules and brain alignment generalize across species? We extend our prior finding that untrained CNNs match backpr

safetyarxiv-cs-lg
23 May 2026
Safety

DecepChain: Inducing Deceptive Reasoning in Large Language Models

DGX agent

arXiv:2510.00319v2 Announce Type: replace Abstract: Large Language Models (LLMs) have been demonstrating strong reasoning capability with their chain-of-thoughts (CoT), which are routinely used by hum

safetyarxiv-cs-lg
23 May 2026
Safety

Don't Forget the Critic: Value-Based Data Rehearsal for Multi-Cyclic Continual Reinforcement Learning

DGX agent

arXiv:2605.22454v1 Announce Type: new Abstract: Data rehearsal has emerged as a leading approach for mitigating catastrophic forgetting in Continual Reinforcement Learning (CRL). However, existing wor

safetyarxiv-cs-lg
23 May 2026
Safety

ECPO: Evidence-Coupled Policy Optimization for Evidence-Certified Candidate Ranking

DGX agent

arXiv:2605.21993v1 Announce Type: cross Abstract: Ranking systems used in decision-support settings should not only order candidates but also expose evidence that can be independently checked. We stud

safetyarxiv-cs-lg
23 May 2026
Safety

Embedding-Based Federated Learning with Runtime Governance for Iron Deficiency Prediction

DGX agent

arXiv:2605.21563v1 Announce Type: new Abstract: Recent reviews find that the vast majority of published healthcare federated learning (FL) studies never reach real-world deployment. We developed an em

safetyarxiv-cs-lg
23 May 2026
Safety

extit{BlockFormer} : Transformer-based inference from interaction maps

DGX agent

arXiv:2605.21617v1 Announce Type: new Abstract: Inference from interaction maps, such as centromere identification from genome-wide chromosome conformation capture techniques -- notably Hi-C -- can be

safetyarxiv-cs-lg
23 May 2026
Safety

Factored Diffusion Policies:Compositionally Generalized Robot Control with a Single Score Network

DGX agent

arXiv:2605.22596v1 Announce Type: new Abstract: Robotic tasks are typically specified by a tuple of factors, such as the object to be grasped, the obstacles to be avoided, the color of the target, and

safetyarxiv-cs-lg
23 May 2026
Safety

From Snapshots to Trajectories: Learning Single-Cell Gene Expression Dynamics via Conditional Flow Matching

DGX agent

arXiv:2605.22340v1 Announce Type: new Abstract: Single-cell RNA sequencing (scRNA-seq) provides high-dimensional profiles of cellular states, enabling data-driven modeling of cellular dynamics over ti

safetyarxiv-cs-lg
23 May 2026
Safety

Generative Modeling by Value-Driven Transport

DGX agent

arXiv:2605.22507v1 Announce Type: new Abstract: We propose a new framework for generative modeling based on a discrete-time stochastic control formulation of measure transport. Adapting classic result

safetyarxiv-cs-lg
23 May 2026
Safety

Harnesses for Inference-Time Alignment over Execution Trajectories

DGX agent

arXiv:2605.21516v1 Announce Type: new Abstract: Harness engineering has emerged as an important inference-time technique for large language model (LLM) agents, aiming to improve long-term performance

safetyarxiv-cs-lg
23 May 2026
Safety

HealthMamba: An Uncertainty-aware Spatiotemporal Graph State Space Model for Effective and Reliable Healthcare Facility Visit Prediction

DGX agent

arXiv:2602.05286v3 Announce Type: replace Abstract: Healthcare facility visit prediction is essential for optimizing healthcare resource allocation and informing public health policy. Despite advanced

safetyarxiv-cs-lg
23 May 2026
Safety

Heterogeneous Agent Collaborative Reinforcement Learning

DGX agent

arXiv:2603.02604v2 Announce Type: replace Abstract: We introduce Heterogeneous Agent Collaborative Reinforcement Learning (HACRL), a new Reinforcement Learning from Verifiable Reward (RLVR) problem th

safetyarxiv-cs-lg
23 May 2026
Safety

Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators

DGX agent

arXiv:2605.22717v1 Announce Type: cross Abstract: Interactive streaming music generation promises the use of generative models for live performance and co-creation that is impossible with offline mode

safetyarxiv-cs-lg
23 May 2026
← Previous
1…164165166167168…260
Next →