AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Quantitative Video World Model Evaluation for Geometric-Consistency

DGX agent

arXiv:2605.15185v1 Announce Type: cross Abstract: Generative video models are increasingly studied as implicit world models, yet evaluating whether they produce physically plausible 3D structure and m

safetyarxiv-cs-ai
15 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability

DGX agent

arXiv:2511.17367v2 Announce Type: replace Abstract: Computing worst-case robust strategies in pursuit-evasion games (PEGs) is time-consuming, especially when real-world factors like partial observabil

safetyarxiv-cs-lg
15 May 2026
Safety

R2R2: Robust Representation for Intensive Experience Reuse via Redundancy Reduction in Self-Predictive Learning

DGX agent

arXiv:2605.14026v1 Announce Type: cross Abstract: For reinforcement learning in data-scarce domains like real-world robotics, intensive data reuse enhances efficiency but induces overfitting. While pr

safetyarxiv-cs-ai
15 May 2026
Safety

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling

DGX agent

arXiv:2510.20206v2 Announce Type: replace Abstract: Prompt design plays a crucial role in text-to-video (T2V) generation, yet user-provided prompts are often short, unstructured, and misaligned with t

safetyarxiv-cs-cv
15 May 2026
Safety

RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO

DGX agent

arXiv:2605.15190v1 Announce Type: new Abstract: Causal autoregressive video diffusion models support real-time streaming generation by extrapolating future chunks from previously generated content. Di

safetyarxiv-cs-cv
15 May 2026
Safety

Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases

DGX agent

arXiv:2601.03630v2 Announce Type: replace Abstract: This paper presents the first systematic comparison investigating whether Large Reasoning Models (LRMs) are superior judges to non-reasoning LLMs. O

safetyarxiv-cs-cl
15 May 2026
Safety

Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages

DGX agent

arXiv:2603.12554v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has been effective for post-training autoregressive (AR) language models, but extending these methods to diffusion

safetyarxiv-cs-ai
15 May 2026
Safety

Reinforcement Learning with Semantic Rewards Enables Low-Resource Language Expansion without Alignment Tax

DGX agent

arXiv:2605.14366v1 Announce Type: new Abstract: Extending large language models (LLMs) to low-resource languages often incurs an 'alignment tax': improvements in the target language come at the cost o

safetyarxiv-cs-cl
15 May 2026
Safety

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy

DGX agent

arXiv:2605.14558v1 Announce Type: cross Abstract: Agentic reinforcement learning trains large language models using multi-turn trajectories that interleave long reasoning traces with short environment

safetyarxiv-cs-ai
15 May 2026
Safety

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models

DGX agent

arXiv:2512.21651v2 Announce Type: replace Abstract: Large Language Models (LLMs) deliver strong performance across a wide range of NLP tasks, but their massive sizes hinder deployment on resource-cons

safetyarxiv-cs-lg
15 May 2026
Safety

REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding

DGX agent

arXiv:2511.13026v3 Announce Type: replace Abstract: Self-reflection mechanisms that rely on purely text-based rethinking processes perform well in most multimodal tasks. However, when directly applied

safetyarxiv-cs-cv
15 May 2026
Safety

ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization

DGX agent

arXiv:2605.14497v1 Announce Type: cross Abstract: Offline-to-online reinforcement learning harnesses the stability of offline pretraining and the flexibility of online fine-tuning. A key challenge lie

safetyarxiv-cs-ai
15 May 2026
Safety

Road Maps as Free Geometric Priors: Weather-Invariant Drone Geo-Localization with GeoFuse

DGX agent

arXiv:2605.14925v1 Announce Type: new Abstract: Drone-view geo-localization aims to match a query drone image, often captured under adverse weather conditions (e.g., rain, snow, fog), against a galler

safetyarxiv-cs-cv
15 May 2026
Safety

Second-Order Actor-Critic Methods for Discounted MDPs via Policy Hessian Decomposition

DGX agent

arXiv:2605.14982v1 Announce Type: cross Abstract: We address the discounted reward setting in reinforcement learning (RL). To mitigate the value approximation challenges in policy gradient methods, ac

safetyarxiv-cs-ai
15 May 2026
Safety

Self-Distilled Agentic Reinforcement Learning

DGX agent

arXiv:2605.15155v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as a central paradigm for post-training LLM agents, yet its trajectory-level reward signal provides only coars

safetyarxiv-cs-ai
15 May 2026
Safety

SimPersona: Learning Discrete Buyer Personas from Raw Clickstreams for Grounded E-Commerce Agents

DGX agent

arXiv:2605.14205v1 Announce Type: new Abstract: LLM-based web agents can navigate live storefronts, yet they often collapse to a single 'average buyer' policy, failing to capture the heterogeneous and

safetyarxiv-cs-ai
15 May 2026
Safety

SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration

DGX agent

arXiv:2605.14089v1 Announce Type: new Abstract: In recent years, a variety of powerful LLM-based agentic systems have been applied to automate complex tasks through task orchestration. However, existi

safetyarxiv-cs-ai
15 May 2026
Safety

Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations

DGX agent

arXiv:2605.14937v1 Announce Type: cross Abstract: Predictive world models enable agents to model scene dynamics and reason about the consequences of their actions. Inspired by human perception, object

safetyarxiv-cs-ai
15 May 2026
Safety

Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings

DGX agent

arXiv:2605.14284v1 Announce Type: new Abstract: Comparative evaluation of multiple dynamic treatment policies is essential for healthcare and policy decisions, yet conventional longitudinal causal inf

safetyarxiv-cs-lg
15 May 2026
Safety

SOCC-ICP: Semantics-Assisted Odometry based on Occupancy Grids and ICP

DGX agent

arXiv:2605.15074v1 Announce Type: new Abstract: Reliable pose estimation in previously unseen environments is a fundamental capability of autonomous systems. Existing LiDAR odometry methods typically

safetyarxiv-cs-ro
15 May 2026
Safety

SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

DGX agent

arXiv:2605.15044v1 Announce Type: cross Abstract: As audio-first agents become increasingly common in physical AI, conversational robots, and screenless wearables, audio large language models (audio-L

safetyarxiv-cs-ai
15 May 2026
Safety

SpectraFlow: Unifying Structural Pretraining and Frequency Adaptation for Medical Image Segmentation

DGX agent

arXiv:2605.14566v1 Announce Type: new Abstract: Medical image segmentation remains challenging in low-data regimes, where scarce annotations often yield poor generalization and ambiguous boundaries wi

safetyarxiv-cs-cv
15 May 2026
Safety

SuperF: Neural Implicit Fields for Multi-Image Super-Resolution

DGX agent

arXiv:2512.09115v2 Announce Type: replace Abstract: High-resolution imagery is often hindered by limitations in sensor technology, atmospheric conditions, and costs. Such challenges occur in satellite

safetyarxiv-cs-cv
15 May 2026
Safety

Temporal Fair Division in Multi-Agent Systems: From Precise Alternation Metrics to Scalable Coordination Proxies

DGX agent

arXiv:2605.14879v1 Announce Type: cross Abstract: A plethora real-world environments require agents to compete repeatedly for the same limited resource, calling for a temporal notion of fairness judge

safetyarxiv-cs-lg
15 May 2026
Safety

TERMS-Bench: Diagnosing LLM Negotiation Agents Beyond Deal Rate

DGX agent

arXiv:2605.13909v1 Announce Type: cross Abstract: Negotiation is a central mechanism of economic exchange, shaping markets, procurement, labor agreements, and resource allocation. It is also a canonic

safetyarxiv-cs-ai
15 May 2026
Safety

The Great Pretender: A Stochasticity Problem in LLM Jailbreak

DGX agent

arXiv:2605.14418v1 Announce Type: cross Abstract: 'Oh-Oh, yes, I'm the great pretender. Pretending that I'm doing well. My need is such, I pretend too much...' summarizes the state in the area of jail

safetyarxiv-cs-ai
15 May 2026
Safety

The Pitfalls of KV Cache Compression

DGX agent

arXiv:2510.00231v2 Announce Type: replace-cross Abstract: KV cache compression promises increased throughput and efficiency with negligible loss in performance. While the gains in throughput are indis

safetyarxiv-cs-ai
15 May 2026
Safety

Towards Continuous Sign Language Conversation from Isolated Signs

DGX agent

arXiv:2605.14705v1 Announce Type: new Abstract: Sign language is the primary language for many Deaf and Hard-of-Hearing (DHH) signers, yet most conversational AI systems still mediate interaction thro

safetyarxiv-cs-cv
15 May 2026
Safety

Training-Free Generative Sampling via Moment-Matched Score Smoothing

DGX agent

arXiv:2605.14276v1 Announce Type: cross Abstract: Diffusion models generate samples by denoising along the score of a perturbed target distribution. In practice, one trains a neural diffusion model, w

safetyarxiv-cs-lg
15 May 2026
Safety

UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars

DGX agent

arXiv:2605.14731v1 Announce Type: cross Abstract: Speech-driven gestures and facial animations are fundamental to expressive digital avatars in games, virtual production, and interactive media. Howeve

safetyarxiv-cs-cv
15 May 2026
Safety

Unbiased and Second-Order-Free Training for High-Dimensional PDEs

DGX agent

arXiv:2605.14643v1 Announce Type: new Abstract: Deep learning methods based on backward stochastic differential equations (BSDEs) have emerged as competitive alternatives to physics-informed neural ne

safetyarxiv-cs-lg
15 May 2026
Safety

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation

DGX agent

arXiv:2605.14626v1 Announce Type: new Abstract: RGB-T semantic segmentation requires strictly aligned VIS-IR-Label triplets; however, such aligned triplet data are often scarce in real-world scenarios

safetyarxiv-cs-cv
15 May 2026
Safety

V2M-Zero: Zero-Pair Time-Aligned Video-to-Music Generation

DGX agent

arXiv:2603.11042v2 Announce Type: replace-cross Abstract: Generating music that temporally aligns with video events is challenging for existing text-to-music models, which lack fine-grained temporal c

safetyarxiv-cs-ai
15 May 2026
Safety

VGGT-Omega

DGX agent

arXiv:2605.15195v1 Announce Type: new Abstract: Recent feed-forward reconstruction models, such as VGGT, have proven competitive with traditional optimization-based reconstructors while also providing

safetyarxiv-cs-cv
15 May 2026
Safety

Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation

DGX agent

arXiv:2602.02994v2 Announce Type: replace Abstract: Reinforcement learning has emerged as a principled post-training paradigm for Temporal Video Grounding (TVG) due to its on-policy optimization, yet

safetyarxiv-cs-cv
15 May 2026
Safety

Vision-Core Guided Contrastive Learning for Balanced Multi-modal Prognosis Prediction of Stroke

DGX agent

arXiv:2605.14710v1 Announce Type: cross Abstract: Deep learning and multi-modal fusion have demonstrated transformative potential in medical diagnosis by integrating diverse data sources. However, acc

safetyarxiv-cs-ai
15 May 2026
Safety

Vision-LLMs for Spatiotemporal Traffic Forecasting

DGX agent

arXiv:2510.11282v2 Announce Type: replace Abstract: Accurate spatiotemporal traffic forecasting is a critical prerequisite for proactive resource management in dense urban mobile networks. While large

safetyarxiv-cs-lg
15 May 2026
Safety

Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video

DGX agent

arXiv:2605.15182v1 Announce Type: new Abstract: Camera-controlled video generation has made substantial progress, enabling generated videos to follow prescribed viewpoint trajectories. However, existi

safetyarxiv-cs-cv
15 May 2026
Safety

XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations

DGX agent

arXiv:2511.02776v2 Announce Type: replace Abstract: Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. Howe

safetyarxiv-cs-ro
15 May 2026
Safety

A_3B_2: Adaptive Asymmetric Adapter for Alleviating Branch Bias in Vision-Language Image Classification with Few-Shot Learning

DGX agent

arXiv:2605.13161v1 Announce Type: new Abstract: Efficient transfer learning methods for large-scale vision-language models (e.g., CLIP) enable strong few-shot transfer, yet existing adaptation methods

safetyarxiv-cs-cv
14 May 2026
Safety

Achieving epsilon^{-2} Sample Complexity for Single-Loop Actor-Critic under Minimal Assumptions

DGX agent

arXiv:2605.13639v1 Announce Type: new Abstract: In this paper, we establish last-iterate convergence rates for off-policy actor--critic methods in reinforcement learning. In particular, under a single

safetyarxiv-cs-lg
14 May 2026
Safety

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

DGX agent

arXiv:2605.12569v1 Announce Type: cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environ

safetyarxiv-cs-ai
14 May 2026
Safety

Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization

DGX agent

arXiv:2605.12878v1 Announce Type: cross Abstract: We propose Adam-SHANG, a Lyapunov-guided Adam-type method that couples momentum, adaptive preconditioning, and a curvature-aware correction through a

safetyarxiv-cs-lg
14 May 2026
Safety

Adaptive mine planning under geological uncertainty: A POMDP framework for sequential decision-making

DGX agent

arXiv:2605.13702v1 Announce Type: new Abstract: Strategic mine production scheduling under geological uncertainty is conventionally formulated as a stochastic optimization problem in which a fixed ext

safetyarxiv-cs-ai
14 May 2026
Safety

Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization

DGX agent

arXiv:2605.12771v1 Announce Type: cross Abstract: Multi-objective reinforcement learning in robotic domains requires balancing complex, non-convex trade-offs between conflicting objectives. While line

safetyarxiv-cs-ai
14 May 2026
Safety

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

DGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

safetyarxiv-cs-lg
14 May 2026
Safety

Addressing Finite-Horizon MDPs via Low-Rank Tensor Value Approximation

DGX agent

arXiv:2501.10598v3 Announce Type: replace Abstract: We study the problem of learning optimal policies in finite-horizon Markov Decision Processes (MDPs) using low-rank reinforcement learning (RL) meth

safetyarxiv-cs-lg
14 May 2026
Safety

Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding

DGX agent

arXiv:2602.02977v2 Announce Type: replace-cross Abstract: Vision-language models such as CLIP often struggle to faithfully understand long, detail-rich captions, relying on dominant scene cues while o

safetyarxiv-cs-ai
14 May 2026
← Previous
1…181182183184185…260
Next →