AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
15 May 2026

Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases

SafetyDGX agent

arXiv:2601.03630v2 Announce Type: replace Abstract: This paper presents the first systematic comparison investigating whether Large Reasoning Models (LRMs) are superior judges to non-reasoning LLMs. O

Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages

SafetyDGX agent

arXiv:2603.12554v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has been effective for post-training autoregressive (AR) language models, but extending these methods to diffusion

Reinforcement Learning with Semantic Rewards Enables Low-Resource Language Expansion without Alignment Tax

SafetyDGX agent

arXiv:2605.14366v1 Announce Type: new Abstract: Extending large language models (LLMs) to low-resource languages often incurs an 'alignment tax': improvements in the target language come at the cost o


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy

SafetyDGX agent

arXiv:2605.14558v1 Announce Type: cross Abstract: Agentic reinforcement learning trains large language models using multi-turn trajectories that interleave long reasoning traces with short environment

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models

SafetyDGX agent

arXiv:2512.21651v2 Announce Type: replace Abstract: Large Language Models (LLMs) deliver strong performance across a wide range of NLP tasks, but their massive sizes hinder deployment on resource-cons

REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding

SafetyDGX agent

arXiv:2511.13026v3 Announce Type: replace Abstract: Self-reflection mechanisms that rely on purely text-based rethinking processes perform well in most multimodal tasks. However, when directly applied

ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization

SafetyDGX agent

arXiv:2605.14497v1 Announce Type: cross Abstract: Offline-to-online reinforcement learning harnesses the stability of offline pretraining and the flexibility of online fine-tuning. A key challenge lie

Road Maps as Free Geometric Priors: Weather-Invariant Drone Geo-Localization with GeoFuse

SafetyDGX agent

arXiv:2605.14925v1 Announce Type: new Abstract: Drone-view geo-localization aims to match a query drone image, often captured under adverse weather conditions (e.g., rain, snow, fog), against a galler

Safety-Constrained Reinforcement Learning with Post-Training Reachability Verification for Robot Navigation

SafetyDGX agent

arXiv:2605.14174v1 Announce Type: new Abstract: Safe navigation for mobile robots demands policies that remain reliable under the high-consequence perception uncertainty of cluttered environments. Yet

Second-Order Actor-Critic Methods for Discounted MDPs via Policy Hessian Decomposition

SafetyDGX agent

arXiv:2605.14982v1 Announce Type: cross Abstract: We address the discounted reward setting in reinforcement learning (RL). To mitigate the value approximation challenges in policy gradient methods, ac

Selective Safety Steering via Value-Filtered Decoding

SafetyDGX agent

arXiv:2605.14746v1 Announce Type: new Abstract: While large language models (LLMs) are trained to align with human values, their generations may still violate safety constraints. A growing line of wor

Self-Distilled Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.15155v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as a central paradigm for post-training LLM agents, yet its trajectory-level reward signal provides only coars

Send the arXiv AI-generated slop, get a yearlong vacation from submissions

SafetyDGX agent

ArXiv will ban authors for one year if they submit papers containing obviously AI-generated content , with examples including hallucinated citations, placeholder text, or chatbot meta-comments left in

SimPersona: Learning Discrete Buyer Personas from Raw Clickstreams for Grounded E-Commerce Agents

SafetyDGX agent

arXiv:2605.14205v1 Announce Type: new Abstract: LLM-based web agents can navigate live storefronts, yet they often collapse to a single 'average buyer' policy, failing to capture the heterogeneous and

SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration

SafetyDGX agent

arXiv:2605.14089v1 Announce Type: new Abstract: In recent years, a variety of powerful LLM-based agentic systems have been applied to automate complex tasks through task orchestration. However, existi

Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations

SafetyDGX agent

arXiv:2605.14937v1 Announce Type: cross Abstract: Predictive world models enable agents to model scene dynamics and reason about the consequences of their actions. Inspired by human perception, object

Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings

SafetyDGX agent

arXiv:2605.14284v1 Announce Type: new Abstract: Comparative evaluation of multiple dynamic treatment policies is essential for healthcare and policy decisions, yet conventional longitudinal causal inf

SOCC-ICP: Semantics-Assisted Odometry based on Occupancy Grids and ICP

SafetyDGX agent

arXiv:2605.15074v1 Announce Type: new Abstract: Reliable pose estimation in previously unseen environments is a fundamental capability of autonomous systems. Existing LiDAR odometry methods typically

SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

SafetyDGX agent

arXiv:2605.15044v1 Announce Type: cross Abstract: As audio-first agents become increasingly common in physical AI, conversational robots, and screenless wearables, audio large language models (audio-L

SpectraFlow: Unifying Structural Pretraining and Frequency Adaptation for Medical Image Segmentation

SafetyDGX agent

arXiv:2605.14566v1 Announce Type: new Abstract: Medical image segmentation remains challenging in low-data regimes, where scarce annotations often yield poor generalization and ambiguous boundaries wi

strawmanning Gary Marcus should be an Olympic sport, there are so many entrants. Hinton wins gold, for apparently faking a quote and putting…

SafetyDGX agent

Gary Marcus criticizes Geoffrey Hinton for allegedly misrepresenting or fabricating a quote attributed to him in a debate about AI. The post sarcastically compares the frequency of strawmanning argume

SuperF: Neural Implicit Fields for Multi-Image Super-Resolution

SafetyDGX agent

arXiv:2512.09115v2 Announce Type: replace Abstract: High-resolution imagery is often hindered by limitations in sensor technology, atmospheric conditions, and costs. Such challenges occur in satellite

Synthesizing POMDP Policies: Sampling Meets Model-checking via Learning

SafetyDGX agent

arXiv:2605.14440v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are the standard framework for decision-making under uncertainty. While sampling-based methods s

Systematic Discovery of Semantic Attacks in Online Map Construction through Conditional Diffusion

SafetyDGX agent

arXiv:2605.14396v1 Announce Type: new Abstract: Autonomous vehicles depend on online HD map construction to perceive lane boundaries, dividers, and pedestrian crossings -- safety-critical road element

Temporal Fair Division in Multi-Agent Systems: From Precise Alternation Metrics to Scalable Coordination Proxies

SafetyDGX agent

arXiv:2605.14879v1 Announce Type: cross Abstract: A plethora real-world environments require agents to compete repeatedly for the same limited resource, calling for a temporal notion of fairness judge

TERMS-Bench: Diagnosing LLM Negotiation Agents Beyond Deal Rate

SafetyDGX agent

arXiv:2605.13909v1 Announce Type: cross Abstract: Negotiation is a central mechanism of economic exchange, shaping markets, procurement, labor agreements, and resource allocation. It is also a canonic

The Great Pretender: A Stochasticity Problem in LLM Jailbreak

SafetyDGX agent

arXiv:2605.14418v1 Announce Type: cross Abstract: 'Oh-Oh, yes, I'm the great pretender. Pretending that I'm doing well. My need is such, I pretend too much...' summarizes the state in the area of jail

The Pitfalls of KV Cache Compression

SafetyDGX agent

arXiv:2510.00231v2 Announce Type: replace-cross Abstract: KV cache compression promises increased throughput and efficiency with negligible loss in performance. While the gains in throughput are indis

Towards Continuous Sign Language Conversation from Isolated Signs

SafetyDGX agent

arXiv:2605.14705v1 Announce Type: new Abstract: Sign language is the primary language for many Deaf and Hard-of-Hearing (DHH) signers, yet most conversational AI systems still mediate interaction thro

Training-Free Generative Sampling via Moment-Matched Score Smoothing

SafetyDGX agent

arXiv:2605.14276v1 Announce Type: cross Abstract: Diffusion models generate samples by denoising along the score of a perturbed target distribution. In practice, one trains a neural diffusion model, w

Training ML Models with Predictable Failures

SafetyDGX agent

arXiv:2605.15134v1 Announce Type: new Abstract: Estimating how often an ML model will fail at deployment scale is central to pre-deployment safety assessment, but a feasible evaluation set is rarely l

UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars

SafetyDGX agent

arXiv:2605.14731v1 Announce Type: cross Abstract: Speech-driven gestures and facial animations are fundamental to expressive digital avatars in games, virtual production, and interactive media. Howeve

Unbiased and Second-Order-Free Training for High-Dimensional PDEs

SafetyDGX agent

arXiv:2605.14643v1 Announce Type: new Abstract: Deep learning methods based on backward stochastic differential equations (BSDEs) have emerged as competitive alternatives to physics-informed neural ne

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation

SafetyDGX agent

arXiv:2605.14626v1 Announce Type: new Abstract: RGB-T semantic segmentation requires strictly aligned VIS-IR-Label triplets; however, such aligned triplet data are often scarce in real-world scenarios

V2M-Zero: Zero-Pair Time-Aligned Video-to-Music Generation

SafetyDGX agent

arXiv:2603.11042v2 Announce Type: replace-cross Abstract: Generating music that temporally aligns with video events is challenging for existing text-to-music models, which lack fine-grained temporal c

Very thoughtful post on the role and success of open source, including potential applications for AVs and AI. Long, but worth the read. -- '…

SafetyDGX agent

Very thoughtful post on the role and success of open source, including potential applications for AVs and AI. Long, but worth the read. -- 'Open source is no longer just how good software gets built.

VGGT-Omega

SafetyDGX agent

arXiv:2605.15195v1 Announce Type: new Abstract: Recent feed-forward reconstruction models, such as VGGT, have proven competitive with traditional optimization-based reconstructors while also providing

Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation

SafetyDGX agent

arXiv:2602.02994v2 Announce Type: replace Abstract: Reinforcement learning has emerged as a principled post-training paradigm for Temporal Video Grounding (TVG) due to its on-policy optimization, yet

Vision-Core Guided Contrastive Learning for Balanced Multi-modal Prognosis Prediction of Stroke

SafetyDGX agent

arXiv:2605.14710v1 Announce Type: cross Abstract: Deep learning and multi-modal fusion have demonstrated transformative potential in medical diagnosis by integrating diverse data sources. However, acc

Vision-LLMs for Spatiotemporal Traffic Forecasting

SafetyDGX agent

arXiv:2510.11282v2 Announce Type: replace Abstract: Accurate spatiotemporal traffic forecasting is a critical prerequisite for proactive resource management in dense urban mobile networks. While large

Viverra: Text-to-Code with Guarantees

SafetyDGX agent

arXiv:2605.14972v1 Announce Type: cross Abstract: A fundamental limitation of Text-to-Code is that no guarantee can be obtained about the correctness of the generated code. Therefore, to ensure its co

Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video

SafetyDGX agent

arXiv:2605.15182v1 Announce Type: new Abstract: Camera-controlled video generation has made substantial progress, enabling generated videos to follow prescribed viewpoint trajectories. However, existi

“we are below rock bottom now”. Trump admin continues to piss away America’s scientific lead for no good reason. @elonmusk stands by, does n…

SafetyDGX agent

“we are below rock bottom now”. Trump admin continues to piss away America’s scientific lead for no good reason. @elonmusk stands by, does nothing. NEW: A serious staffing shortage — of the Trump admi

XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations

SafetyDGX agent

arXiv:2511.02776v2 Announce Type: replace Abstract: Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. Howe

14 May 2026

140,000 fake citations in 2025 alone😠

SafetyDGX agent

Gary Marcus highlighted a concerning problem in 2025 where approximately 140,000 fabricated citations were identified, likely referring to false or hallucinated references generated by AI systems in a

A Data Efficiency Study of Synthetic Fog for Object Detection Using the Clear2Fog Pipeline

SafetyDGX agent

arXiv:2605.12608v1 Announce Type: new Abstract: Object detection in adverse weather is critical for the safety of autonomous vehicles; however, the scarcity of labelled, real-world foggy data remains

A Five-Layer MLOps Architecture for Connected Automated Driving

SafetyDGX agent

arXiv:2605.12719v1 Announce Type: cross Abstract: The continual assurance of safety and performance of automated driving systems (ADSs) poses significant challenges. ADSs operate in complex, dynamic,

A_3B_2: Adaptive Asymmetric Adapter for Alleviating Branch Bias in Vision-Language Image Classification with Few-Shot Learning

SafetyDGX agent

arXiv:2605.13161v1 Announce Type: new Abstract: Efficient transfer learning methods for large-scale vision-language models (e.g., CLIP) enable strong few-shot transfer, yet existing adaptation methods

Achieving epsilon^{-2} Sample Complexity for Single-Loop Actor-Critic under Minimal Assumptions

SafetyDGX agent

arXiv:2605.13639v1 Announce Type: new Abstract: In this paper, we establish last-iterate convergence rates for off-policy actor--critic methods in reinforcement learning. In particular, under a single

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

SafetyDGX agent

arXiv:2605.12569v1 Announce Type: cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environ

Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization

SafetyDGX agent

arXiv:2605.12878v1 Announce Type: cross Abstract: We propose Adam-SHANG, a Lyapunov-guided Adam-type method that couples momentum, adaptive preconditioning, and a curvature-aware correction through a

Adaptive Conformal Prediction for Reliable and Explainable Medical Image Classification

SafetyDGX agent

arXiv:2605.12917v1 Announce Type: new Abstract: Deep learning models for medical imaging often exhibit overconfidence, creating safety risks in ambiguous diagnostic scenarios. While Conformal Predicti

Adaptive mine planning under geological uncertainty: A POMDP framework for sequential decision-making

SafetyDGX agent

arXiv:2605.13702v1 Announce Type: new Abstract: Strategic mine production scheduling under geological uncertainty is conventionally formulated as a stochastic optimization problem in which a fixed ext

Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization

SafetyDGX agent

arXiv:2605.12771v1 Announce Type: cross Abstract: Multi-objective reinforcement learning in robotic domains requires balancing complex, non-convex trade-offs between conflicting objectives. While line

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

SafetyDGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

Addressing Finite-Horizon MDPs via Low-Rank Tensor Value Approximation

SafetyDGX agent

arXiv:2501.10598v3 Announce Type: replace Abstract: We study the problem of learning optimal policies in finite-horizon Markov Decision Processes (MDPs) using low-rank reinforcement learning (RL) meth

AgenticAITA: A Proof-Of-Concept About Deliberative Multi-Agent Reasoning for Autonomous Trading Systems

SafetyDGX agent

arXiv:2605.12532v1 Announce Type: cross Abstract: Conventional algorithmic trading systems are grounded in deterministic heuristics or offline-trained statistical models that cannot adapt to the seman

AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions

SafetyDGX agent

arXiv:2408.12935v4 Announce Type: replace Abstract: AI Safety is an emerging area of critical importance to the safe adoption and deployment of AI systems. With the rapid proliferation of AI and espec

Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding

SafetyDGX agent

arXiv:2602.02977v2 Announce Type: replace-cross Abstract: Vision-language models such as CLIP often struggle to faithfully understand long, detail-rich captions, relying on dominant scene cues while o

Aligning Network Equivariance with Data Symmetry: A Theoretical Framework and Adaptive Approach for Image Restoration

SafetyDGX agent

arXiv:2605.13744v1 Announce Type: new Abstract: Image restoration is an inherently ill posed inverse problem. Equivariant networks that embed geometric symmetry priors can mitigate this ill posedness

← Previous
1…141142143144145…214
Next →