AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs

DGX agent

arXiv:2605.15621v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong multimodal understanding, but their inference cost grows rapidly with the number of visual tokens, e

safetyarxiv-cs-cv
18 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LUIVITON: Learned Universal Interoperable VIrtual Try-ON

DGX agent

arXiv:2509.05030v2 Announce Type: replace Abstract: To enable large-scale reuse of real-world 3D assets, where garments and characters rarely share skeletons, templates, or dense correspondences, we p

safetyarxiv-cs-cv
18 May 2026
Safety

MAgSeg: Segmentation of Agricultural Landscapes in High-Resolution Satellite Imagery using Multimodal Large Language Models

DGX agent

arXiv:2605.16179v1 Announce Type: new Abstract: Agricultural landscape segmentation in the Global South is challenging as it is characterized by fragmented plots, high intra-class variance, and a scar

safetyarxiv-cs-cv
18 May 2026
Safety

MaTe: Images Are All You Need for Material Transfer via Diffusion Transformer

DGX agent

arXiv:2605.15660v1 Announce Type: new Abstract: Recent diffusion-based methods for material transfer rely on image fine-tuning or complex architectures with assistive networks, but face challenges inc

safetyarxiv-cs-cv
18 May 2026
Safety

Mind Dreamer: Untethering Imagination via Active Latent Intervention on Latent Manifolds

DGX agent

arXiv:2605.16030v1 Announce Type: new Abstract: Model-Based Reinforcement Learning (MBRL) leverages latent imagination for sample efficiency, yet remains constrained by Historical Tethering: imaginati

safetyarxiv-cs-lg
18 May 2026
Safety

Missing-Data-Induced Phase Transitions in Spectral PLS for Multimodal Learning

DGX agent

arXiv:2601.21294v2 Announce Type: replace Abstract: Partial Least Squares (PLS) learns shared structure from paired data via the top singular vectors of the empirical cross-covariance (PLS-SVD), but m

safetyarxiv-cs-lg
18 May 2026
Safety

Monotone and Separable Set Functions: Characterizations and Neural Models

DGX agent

arXiv:2510.23634v3 Announce Type: replace-cross Abstract: Motivated by applications for set containment problems, we consider the following fundamental problem: can we design set-to-vector functions s

safetyarxiv-cs-ai
18 May 2026
Safety

NavRL++: A System-Level Framework for Improving Sim-to-Real Transfer in Reinforcement Learning-Based Robot Navigation

DGX agent

arXiv:2605.15559v1 Announce Type: new Abstract: Recent years have witnessed significant progress in autonomous navigation using reinforcement learning. However, existing approaches largely emphasize r

safetyarxiv-cs-ro
18 May 2026
Safety

Neutral-Reference Prompting for Vision-Language Models

DGX agent

arXiv:2605.15615v1 Announce Type: new Abstract: Efficient transfer learning of vision-language models (VLMs) commonly suffers from a Base-New Trade-off (BNT): improving performance on unseen (new) cla

safetyarxiv-cs-cv
18 May 2026
Safety

Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR

DGX agent

arXiv:2605.15726v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a scalable paradigm for improving the reasoning capabilities of large language mode

safetyarxiv-cs-ai
18 May 2026
Safety

Offline Reinforcement Learning with Universal Horizon Models

DGX agent

arXiv:2605.15603v1 Announce Type: cross Abstract: Model-based reinforcement learning (RL) offers a compelling approach to offline RL by enabling value learning on imagined on-policy trajectories. Howe

safetyarxiv-cs-ai
18 May 2026
Safety

OmniVL-Guard: Towards Unified Vision-Language Forgery Detection and Grounding via Balanced RL

DGX agent

arXiv:2602.10687v3 Announce Type: replace-cross Abstract: Existing forgery detection methods are often limited to uni-modal or bi-modal settings, failing to handle the interleaved text, images, and vi

safetyarxiv-cs-ai
18 May 2026
Safety

On Kernel Eigen-alignments of KRR: Reconstruction and Generalization

DGX agent

arXiv:2605.15240v1 Announce Type: cross Abstract: This paper investigates the critical role of eigenalignments between the kernel matrix and learning targets in achieving robust generalization in lear

safetyarxiv-cs-lg
18 May 2026
Safety

OpenFrontier: General Navigation with Visual-Language Grounded Frontiers

DGX agent

arXiv:2603.05377v2 Announce Type: replace-cross Abstract: Open-world navigation requires robots to make decisions in complex everyday environments while adapting to flexible task requirements. Convent

safetyarxiv-cs-cv
18 May 2026
Safety

Opponent State Inference Under Partial Observability: An HMM-POMDP Framework for 2026 Formula 1 Energy Strategy

DGX agent

arXiv:2603.01290v3 Announce Type: replace Abstract: The 2026 Formula 1 technical regulations introduce a fundamental change to energy strategy: under a 50/50 internal combustion engine / battery power

safetyarxiv-cs-ai
18 May 2026
Safety

Pessimistic Risk-Aware Policy Learning in Contextual Bandits

DGX agent

arXiv:2605.15620v1 Announce Type: cross Abstract: We study risk-aware offline policy learning, aiming to learn a decision rule from logged data that is optimal under general risk criteria. This proble

safetyarxiv-cs-lg
18 May 2026
Safety

phi-Balancing for Mixture-of-Experts Training

DGX agent

arXiv:2605.15403v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models rely on balanced expert utilization to fully realize their scalability. However, existing load-balancing methods are lar

safetyarxiv-cs-lg
18 May 2026
Safety

Polynomial Neural Sheaf Diffusion: A Spectral Filtering Approach on Cellular Sheaves

DGX agent

arXiv:2512.00242v3 Announce Type: replace-cross Abstract: Sheaf Neural Networks equip graph structures with a cellular sheaf: a geometric structure which assigns local vector spaces (stalks) and a lin

safetyarxiv-cs-ai
18 May 2026
Safety

Preconditioned Regularized Wasserstein Proximal Sampling

DGX agent

arXiv:2509.01685v2 Announce Type: replace-cross Abstract: We consider sampling from a Gibbs distribution by evolving finitely many particles. We propose a preconditioned version of a recently proposed

safetyarxiv-cs-lg
18 May 2026
Safety

PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding

DGX agent

arXiv:2605.15609v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising masked token sequences. Although dLLMs can predict all masked positions i

safetyarxiv-cs-cl
18 May 2026
Safety

RanSOM: Second-Order Momentum with Randomized Scaling for Constrained and Unconstrained Optimization

DGX agent

arXiv:2602.06824v2 Announce Type: replace-cross Abstract: Momentum methods, such as Polyak's Heavy Ball, are the standard for training deep networks but suffer from curvature-induced bias in stochasti

safetyarxiv-cs-lg
18 May 2026
Safety

RE-SAC: Disentangling aleatoric and epistemic risks in bus fleet control: A stable and robust ensemble DRL approach

DGX agent

arXiv:2603.18396v3 Announce Type: replace Abstract: Bus holding control is challenging due to stochastic traffic and passenger demand. While deep reinforcement learning (DRL) shows promise, standard a

safetyarxiv-cs-lg
18 May 2026
Safety

ReactiveGWM: Steering NPC in Reactive Game World Models

DGX agent

arXiv:2605.15256v1 Announce Type: new Abstract: Current game world models simulate environments from a subjective, player-centric perspective. However, by treating the Non-Player Character (NPC) merel

safetyarxiv-cs-cv
18 May 2026
Safety

ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation

DGX agent

arXiv:2605.16080v1 Announce Type: new Abstract: The rise of AI-generated images (AIGIs) poses growing challenges for digital authenticity, prompting the need for efficient, generalizable image forgery

safetyarxiv-cs-cv
18 May 2026
Safety

Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective

DGX agent

arXiv:2605.15976v1 Announce Type: cross Abstract: Production machine translation relies overwhelmingly on encoder-decoder Seq2Seq models, yet reinforcement learning approaches to MT fine-tuning have l

safetyarxiv-cs-ai
18 May 2026
Safety

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

DGX agent

arXiv:2601.07820v2 Announce Type: replace Abstract: In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers

safetyarxiv-cs-cl
18 May 2026
Safety

Res^2CLIP: Few-Shot Generalist Anomaly Detection with Residual-to-Residual Alignment

DGX agent

arXiv:2605.16171v1 Announce Type: new Abstract: Few-shot Generalist Anomaly Detection requires models to generalize to novel categories without retraining, posing significant challenges in real-world

safetyarxiv-cs-cv
18 May 2026
Safety

Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays

DGX agent

arXiv:2605.15480v1 Announce Type: cross Abstract: Stochastic communication delays in teleoperation introduce signal discontinuities that undermine control stability and degrade control performance. Co

safetyarxiv-cs-ai
18 May 2026
Safety

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

DGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

safetyarxiv-cs-cl
18 May 2026
Safety

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

DGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

safetyarxiv-cs-cv
18 May 2026
Safety

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

DGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

safetyarxiv-cs-ai
18 May 2026
Safety

SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use

DGX agent

arXiv:2601.06366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are transforming enterprise workflows but introduce security and ethics challenges when employees inadvertently s

safetyarxiv-cs-ai
18 May 2026
Safety

ScreenSearch: Uncertainty-Aware OS Exploration

DGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

safetyarxiv-cs-ai
18 May 2026
Safety

Second-Order Multi-Level Variance Correction for Modality Competition in Multimodal Models

DGX agent

arXiv:2605.16165v1 Announce Type: cross Abstract: Autoregressive next-token training offers a unified formulation for image generation and text understanding, but it also creates strong modality compe

safetyarxiv-cs-ai
18 May 2026
Safety

Seeing What Matters: Visual Preference Policy Optimization for Visual Generation

DGX agent

arXiv:2511.18719v4 Announce Type: replace Abstract: Reinforcement learning (RL) has become a powerful tool for post-training visual generative models, with Group Relative Policy Optimization (GRPO) in

safetyarxiv-cs-cv
18 May 2026
Safety

Self-Supervised Learning by Curvature Alignment

DGX agent

arXiv:2511.17426v2 Announce Type: replace-cross Abstract: Self-supervised learning (SSL) has recently advanced through non-contrastive methods that couple an invariance term with variance, covariance,

safetyarxiv-cs-cv
18 May 2026
Safety

Semi-MedRef: Semi-Supervised Medical Referring Image Segmentation with Cross-Modal Alignment

DGX agent

arXiv:2605.15720v1 Announce Type: new Abstract: Medical referring image segmentation (MRIS) requires pixel-level masks aligned with textual descriptions of anatomical locations, making annotation cost

safetyarxiv-cs-cv
18 May 2026
Safety

Sign-Separated Finite-Time Error Analysis of Q-Learning

DGX agent

arXiv:2605.16103v1 Announce Type: new Abstract: This paper develops a sign-separated finite-time error analysis for constant step-size Q-learning. Starting from the switching-system representation, th

safetyarxiv-cs-ai
18 May 2026
Safety

SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

DGX agent

arXiv:2604.02268v2 Announce Type: replace Abstract: Agent skills, structured packages of procedural knowledge and executable resources that agents dynamically load at inference time, have become a rel

safetyarxiv-cs-lg
18 May 2026
Safety

SkiP: When to Skip and When to Refine for Efficient Robot Manipulation

DGX agent

arXiv:2605.15536v1 Announce Type: cross Abstract: Previous imitation learning policies predict future actions at every control step, whether in smooth motion phases or precise, contact-rich operation

safetyarxiv-cs-ai
18 May 2026
Safety

STABLE: Simulation-Ready Tabletop Layout Generation via a Semantics-Physics Dual System

DGX agent

arXiv:2605.16137v1 Announce Type: new Abstract: Generating simulation-ready tabletop scenes from task instructions is an intriguing and promising research direction in the field of Embodied AI. Howeve

safetyarxiv-cs-cv
18 May 2026
Safety

Task-Semantic Graph-Driven Distributed Agent Networking for Underwater Target Tracking

DGX agent

arXiv:2605.15528v1 Announce Type: new Abstract: Autonomous underwater vehicle (AUV) swarms are emerging as intelligent underwater networks, where each node must sense, communicate, process local data,

safetyarxiv-cs-ro
18 May 2026
Safety

TemplateRL: Structured Template-Guided Reinforcement Learning for LLM Reasoning

DGX agent

arXiv:2505.15692v5 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as an effective paradigm for enhancing model reasoning. However, existing RL methods like GRPO typically rel

safetyarxiv-cs-cl
18 May 2026
Safety

Terrain Consistent Reference-Guided RL for Humanoid Navigation Autonomy

DGX agent

arXiv:2605.15517v1 Announce Type: new Abstract: We present a method for training reference-guided, perceptive reinforcement learning locomotion policies for humanoid robots in which reference trajecto

safetyarxiv-cs-ro
18 May 2026
Safety

TopoEvo: A Topology-Aware Self-Evolving Multi-Agent Framework for Root Cause Analysis in Microservices

DGX agent

arXiv:2605.15611v1 Announce Type: new Abstract: Root cause analysis (RCA) in microservices is challenging due to (i) noisy and heterogeneous multimodal observability (metrics, logs, traces), (ii) casc

safetyarxiv-cs-ai
18 May 2026
Safety

Towards Code-Oriented LM Embeddings for Surrogate-Assisted Neural Architecture Search

DGX agent

arXiv:2605.15649v1 Announce Type: new Abstract: Developing effective surrogates (performance predictors) for Neural Architecture Search (NAS) typically requires expensive fine-tuning or the engineerin

safetyarxiv-cs-lg
18 May 2026
Safety

Unified High-Probability Analysis of Stochastic Variance-Reduced Estimation

DGX agent

arXiv:2605.15388v1 Announce Type: new Abstract: Stochastic estimators are fundamental to large-scale optimization, where population quantities must be inferred from noisy oracle observations. Although

safetyarxiv-cs-lg
18 May 2026
Safety

Unveiling the Black Box: A Multi-Layer Framework for Explaining Reinforcement Learning-Based Cyber Agents

DGX agent

arXiv:2505.11708v3 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) agents are increasingly used to simulate sophisticated cyberattacks, but their decision-making processes remain op

safetyarxiv-cs-lg
18 May 2026
← Previous
1…178179180181182…260
Next →