AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Privacy-Aware Video Anomaly Detection through Orthogonal Subspace Projection

DGX agent

arXiv:2605.08651v1 Announce Type: cross Abstract: Video anomaly detection (VAD) systems often prioritize accuracy while overlooking privacy concerns, limiting their suitability for real-world deployme

safetyarxiv-cs-ai
12 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation

DGX agent

arXiv:2605.08774v1 Announce Type: cross Abstract: Long-horizon robotic manipulation requires dense feedback that reflects how a task advances through its procedural stages, not merely whether the fina

safetyarxiv-cs-lg
12 May 2026
Safety

ProteinOPD: Towards Effective and Efficient Preference Alignment for Protein Design

DGX agent

arXiv:2605.10189v1 Announce Type: cross Abstract: Designing proteins with desired functions or properties represents a core goal in synthetic biology and drug discovery. Recent advances in protein lan

safetyarxiv-cs-ai
12 May 2026
Safety

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions

DGX agent

arXiv:2605.09893v1 Announce Type: cross Abstract: Large language models (LLMs) are often evaluated based on their stated values, yet these do not reliably translate into their actions, a discrepancy t

safetyarxiv-cs-ai
12 May 2026
Safety

Q-learning with Adjoint Matching

DGX agent

arXiv:2601.14234v3 Announce Type: replace-cross Abstract: We propose Q-learning with Adjoint Matching (QAM), a novel TD-based reinforcement learning (RL) algorithm that tackles a long-standing challen

safetyarxiv-cs-ai
12 May 2026
Safety

Quantile-Coupled Flow Matching for Distributional Reinforcement Learning

DGX agent

arXiv:2605.08515v1 Announce Type: new Abstract: Unlike standard expected-return Reinforcement Learning (RL), Distributional RL (DRL) models the full return distribution, making it better-suited for un

safetyarxiv-cs-lg
12 May 2026
Safety

Re-Triggering Safeguards within LLMs for Jailbreak Detection

DGX agent

arXiv:2605.10611v1 Announce Type: cross Abstract: This paper proposes a jailbreaking prompt detection method for large language models (LLMs) to defend against jailbreak attacks. Although recent LLMs

safetyarxiv-cs-ai
12 May 2026
Safety

Reasoning Compression with Mixed-Policy Distillation

DGX agent

arXiv:2605.08776v1 Announce Type: new Abstract: Reasoning-centric large language models (LLMs) achieve strong performance by generating intermediate reasoning trajectories, but often incur excessive t

safetyarxiv-cs-ai
12 May 2026
Safety

Reasoning Is Not Free: Robust Adaptive Cost-Efficient Routing for LLM-as-a-Judge

DGX agent

arXiv:2605.10805v1 Announce Type: new Abstract: Reasoning-capable large language models (LLMs) have recently been adopted as automated judges, but their benefits and costs in LLM-as-a-Judge settings r

safetyarxiv-cs-ai
12 May 2026
Safety

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning

DGX agent

arXiv:2605.09614v1 Announce Type: new Abstract: Long chain-of-thought (CoT) reasoning improves large vision--language models, but visual information often fades during generation, limiting long-horizo

safetyarxiv-cs-cv
12 May 2026
Safety

Reflective Prompted Policy Optimization: Trajectory-Grounded Revision and Salience Bias

DGX agent

arXiv:2605.08315v1 Announce Type: new Abstract: Existing LLM-based policy optimizers see only scalar rewards: that a policy scored 0.45, but not whether the agent got stuck in a loop, fell into a hole

safetyarxiv-cs-lg
12 May 2026
Safety

Reinforcement learning for inverse structural design and rapid laser cutting of kirigami prototypes

DGX agent

arXiv:2605.08098v1 Announce Type: new Abstract: Kirigami is an increasingly useful fabrication method to produce shape-programmable metamaterial structures. However, inverse design remains difficult b

safetyarxiv-cs-lg
12 May 2026
Safety

Reinforcement Learning with Action Chunking

DGX agent

arXiv:2507.07969v4 Announce Type: replace-cross Abstract: We present Q-chunking, a simple yet effective recipe for improving reinforcement learning (RL) algorithms for long-horizon, sparse-reward task

safetyarxiv-cs-ai
12 May 2026
Safety

Reinforcing Multimodal Reasoning Against Visual Degradation

DGX agent

arXiv:2605.09262v1 Announce Type: cross Abstract: Reinforcement Learning has significantly advanced the reasoning capabilities of Multimodal Large Language Models (MLLMs), yet the resulting policies r

safetyarxiv-cs-cl
12 May 2026
Safety

Relational reasoning and inductive bias in transformers and large language models

DGX agent

arXiv:2506.04289v3 Announce Type: replace Abstract: Transformer-based models have demonstrated remarkable reasoning abilities, but the mechanisms underlying relational reasoning remain poorly understo

safetyarxiv-cs-lg
12 May 2026
Safety

Relational Retrieval: Leveraging Known-Novel Interactions for Generalized Category Discovery

DGX agent

arXiv:2605.09420v1 Announce Type: cross Abstract: In this study, we tackle Generalized Category Discovery (GCD) via a Relational Retrieval perspective, explicitly coupling labeled and unlabeled data t

safetyarxiv-cs-ai
12 May 2026
Safety

Relations Are Channels: Knowledge Graph Embedding via Kraus Decompositions

DGX agent

arXiv:2605.10317v1 Announce Type: cross Abstract: Knowledge graph embedding (KGE) models typically represent each relation as an operator on entity embeddings. In this work, we identify three structur

safetyarxiv-cs-ai
12 May 2026
Safety

Relative Score Policy Optimization for Diffusion Language Models

DGX agent

arXiv:2605.10218v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) offer a promising route to parallel and efficient text generation, but improving their reasoning ability require

safetyarxiv-cs-cl
12 May 2026
Safety

Remember to Forget: Gated Adaptive Positional Encoding

DGX agent

arXiv:2605.10414v1 Announce Type: new Abstract: Rotary Positional Encoding (RoPE) is widely used in modern large language models. However, when sequences are extended beyond the range seen during trai

safetyarxiv-cs-lg
12 May 2026
Safety

RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models

DGX agent

arXiv:2605.09410v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models remain brittle in long-horizon, contact-rich manipulation because success-only imitation provides little supervisi

safetyarxiv-cs-ai
12 May 2026
Safety

Responsible Benchmarking of Fairness for Automatic Speech Recognition

DGX agent

arXiv:2605.10615v1 Announce Type: new Abstract: Many studies have shown automatic speech processing (ASR) systems have unequal performance across speakergroups (SG's). However, the manner in which suc

safetyarxiv-cs-cl
12 May 2026
Safety

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models

DGX agent

arXiv:2605.08186v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) via entropy minimization (EM) has proven effective for classification tasks, yet its application to generative autoregressi

safetyarxiv-cs-ai
12 May 2026
Safety

Rethinking Loss Reweighting for Imbalance Learning as an Inverse Problem: A Neural Collapse Point of View

DGX agent

arXiv:2605.10047v1 Announce Type: cross Abstract: Loss reweighting is a widely used strategy for long-tailed classification, but existing reweighting strategies often rely on heuristics and rarely def

safetyarxiv-cs-ai
12 May 2026
Safety

Rethinking Ratio-Based Trust Regions for Policy Optimization in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.09212v1 Announce Type: new Abstract: Centralized training with decentralized execution (CTDE) is a standard framework for cooperative multi-agent policy-gradient reinforcement learning, all

safetyarxiv-cs-lg
12 May 2026
Safety

Rethinking RL for LLM Reasoning: It's Sparse Policy Selection, Not Capability Learning

DGX agent

arXiv:2605.06241v2 Announce Type: replace Abstract: Reinforcement learning has become the standard for improving reasoning in large language models, yet evidence increasingly suggests that RL does not

safetyarxiv-cs-cl
12 May 2026
Safety

Revisiting Policy Gradients for Restricted Policy Classes: Escaping Myopic Local Optima with k-step Policy Gradients

DGX agent

arXiv:2605.10909v1 Announce Type: new Abstract: This work revisits standard policy gradient methods used on restricted policy classes, which are known to get stuck in suboptimal critical points. We id

safetyarxiv-cs-lg
12 May 2026
Safety

Revitalizing the Beginning: Avoiding Storage Dependency for Model Merging in Continual Learning

DGX agent

arXiv:2605.08311v1 Announce Type: cross Abstract: Model merging provides a compelling paradigm for integrating specialized expertise into a unified multi-task model, a goal that aligns naturally with

safetyarxiv-cs-cv
12 May 2026
Safety

Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

DGX agent

arXiv:2512.00920v4 Announce Type: replace Abstract: Reliable reward models (RMs) are critical for ensuring the safe alignment of large language models (LLMs). However, current RM evaluation methods fo

safetyarxiv-cs-cl
12 May 2026
Safety

Reward-Conditioned Reinforcement Learning

DGX agent

arXiv:2603.05066v2 Announce Type: replace Abstract: Single-task RL agents are typically trained under a fixed reward function, which limits their robustness to reward misspecification and their abilit

safetyarxiv-cs-lg
12 May 2026
Safety

RigidFormer: Learning Rigid Dynamics using Transformers

DGX agent

arXiv:2605.09196v1 Announce Type: cross Abstract: Learning-based simulation of multi-object rigid-body dynamics remains difficult because contact is discontinuous and errors compound over long horizon

safetyarxiv-cs-ai
12 May 2026
Safety

Route by State, Recover from Trace: STAR with Failure-Aware Markov Routing for Multi-Agent Spatiotemporal Reasoning

DGX agent

arXiv:2605.10057v1 Announce Type: new Abstract: Compositional spatiotemporal reasoning often requires a system to invoke multiple heterogeneous specialists, such as geometric, temporal, topological, a

safetyarxiv-cs-ai
12 May 2026
Safety

RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards

DGX agent

arXiv:2605.10899v1 Announce Type: new Abstract: Training deep research agents, namely systems that plan, search, evaluate evidence, and synthesize long-form reports, pushes reinforcement learning beyo

safetyarxiv-cs-cl
12 May 2026
Safety

RuPLaR : Efficient Latent Compression of LLM Reasoning Chains with Rule-Based Priors From Multi-Step to One-Step

DGX agent

arXiv:2605.09346v1 Announce Type: cross Abstract: The Chain-of-Thought (CoT) paradigm, while enhancing the interpretability of Large Language Models (LLMs), is constrained by the inefficiencies and ex

safetyarxiv-cs-ai
12 May 2026
Safety

SalesSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators

DGX agent

arXiv:2605.08334v1 Announce Type: new Abstract: We present SalesSim, a framework and testbed for evaluating the ability of Multimodal Large Language Models (MLLMs) to simulate realistic, persona-drive

safetyarxiv-cs-cl
12 May 2026
Safety

Sample-Mean Anchored Thompson Sampling for Offline-to-Online Learning with Distribution Shift

DGX agent

arXiv:2605.10289v1 Announce Type: new Abstract: Offline-to-online learning aims to improve online decision-making by leveraging offline logged data. A central challenge in this setting is the distribu

safetyarxiv-cs-lg
12 May 2026
Safety

SARL: Label-Free Reinforcement Learning by Rewarding Reasoning Topology

DGX agent

arXiv:2603.27977v2 Announce Type: replace Abstract: Reinforcement learning is critical to improving large reasoning models, but its success relies heavily on verifiable rewards (RLVR), making it hard

safetyarxiv-cs-ai
12 May 2026
Safety

SceneFactory: GPU-Accelerated Multi-Agent Driving Simulation with Physics-Based Vehicle Dynamics

DGX agent

arXiv:2605.08528v1 Announce Type: cross Abstract: Autonomous-driving simulators typically trade physical fidelity for scalable parallelism. Physics-based platforms such as CARLA and MetaDrive provide

safetyarxiv-cs-ro
12 May 2026
Safety

SCOT: Multi-Source Cross-City Transfer with Optimal-Transport Soft-Correspondence Objective

DGX agent

arXiv:2604.07383v2 Announce Type: replace Abstract: Cross-city transfer improves prediction in label-scarce cities by leveraging labeled data from other cities, but it becomes challenging when cities

safetyarxiv-cs-lg
12 May 2026
Safety

SDFlow: Similarity-Driven Flow Matching for Time Series Generation

DGX agent

arXiv:2605.05736v2 Announce Type: replace Abstract: Vector quantization (VQ) with autoregressive (AR) token modeling is a widely adopted and highly competitive paradigm for time-series generation. How

safetyarxiv-cs-ai
12 May 2026
Safety

Segment Anything with Robust Uncertainty-Accuracy Correlation

DGX agent

arXiv:2605.10603v1 Announce Type: new Abstract: Despite strong zero-shot performance, SAM is unreliable under domain shift due to Mask-level Confidence Confusion (MCC), where a single IoU-based mask s

safetyarxiv-cs-cv
12 May 2026
Safety

Selection of the Best Policy under Fairness Constraints for Subpopulations

DGX agent

arXiv:2605.09945v1 Announce Type: new Abstract: Many high-stakes decisions in health care, public policy, and clinical development require committing to a single policy that will be applied uniformly

safetyarxiv-cs-lg
12 May 2026
Safety

Selection Plateau and a Sparsity-Dependent Hierarchy of Pruning Features

DGX agent

arXiv:2605.09345v1 Announce Type: new Abstract: We identify a Selection Plateau phenomenon in one-shot neural network pruning: all rank-monotone weight scorers converge to identical accuracy at fixed

safetyarxiv-cs-lg
12 May 2026
Safety

Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2605.08874v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation requires adapting image-level vision-language models such as CLIP to dense pixel-level prediction, which is challe

safetyarxiv-cs-cv
12 May 2026
Safety

SGC-RML: A reliable and interpretable longitudinal assessment for PD in real-world DNS

DGX agent

arXiv:2605.08302v1 Announce Type: cross Abstract: Real-world digital Parkinson's disease assessment faces challenges such as heterogeneous modalities, cross-device bias, and incomplete labeling. Exist

safetyarxiv-cs-ai
12 May 2026
Safety

Sink vs. diagonal patterns as mechanisms for attention switch and oversmoothing prevention

DGX agent

arXiv:2605.08453v1 Announce Type: cross Abstract: This paper studies the role of sinks and diagonal patterns as attention switch and anti-oversmoothing mechanisms. We analyze geometric conditions unde

safetyarxiv-cs-ai
12 May 2026
Safety

SKG-VLA: Scene Knowledge Graph Priors for Structured Scene Semantics and Multimodal Reasoning for Decision Making

DGX agent

arXiv:2605.09343v1 Announce Type: new Abstract: Decision making in large-scale complaint handling systems increasingly relies on heterogeneous evidence, including complaint narratives, screenshots, or

safetyarxiv-cs-ai
12 May 2026
Safety

Skill-R1: Agent Skill Evolution via Reinforcement Learning

DGX agent

arXiv:2605.09359v1 Announce Type: cross Abstract: Agentic large language models often rely on skills, reusable natural language procedures that guide planning, action, and tool use. In practice, skill

safetyarxiv-cs-ai
12 May 2026
Safety

SLASH the Sink: Sharpening Structural Attention Inside LLMs

DGX agent

arXiv:2605.10503v1 Announce Type: new Abstract: Large Language Models (LLMs) show remarkable semantic understanding but often struggle with structural understanding when processing graph topologies in

safetyarxiv-cs-ai
12 May 2026
← Previous
1…190191192193194…257
Next →