AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
15 May 2026

Let Robots Feel Your Touch: Visuo-Tactile Cortical Alignment for Embodied Mirror Resonance

SafetyDGX agent

arXiv:2605.14571v1 Announce Type: cross Abstract: Observing touch on another's body can elicit corresponding tactile sensations in the observer, a phenomenon termed mirror touch that supports empathy

Logging Policy Design for Off-Policy Evaluation

SafetyDGX agent

arXiv:2605.15108v1 Announce Type: cross Abstract: Off-policy evaluation (OPE) estimates the value of a target treatment policy (e.g., a recommender system) using data collected by a different logging

LPH-VTON: Resolving the Structure-Texture Dilemma of Virtual Try-On via Latent Process Handover

SafetyDGX agent

arXiv:2605.14874v1 Announce Type: new Abstract: Virtual Try-On (VTON) aims to synthesize photorealistic images of garments precisely aligned with a person's body and pose. Current diffusion-based meth

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

SafetyDGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

Medical Report Generation: A Hierarchical Task Structure-Based Cross-Modal Causal Intervention Framework

SafetyDGX agent

arXiv:2511.02271v2 Announce Type: replace Abstract: Medical Report Generation (MRG) is a key part of modern medical diagnostics, as it automatically generates reports from radiological images to reduc

Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models

SafetyDGX agent

arXiv:2605.14530v1 Announce Type: new Abstract: Large diffusion vision-language models (LDVLMs) have recently emerged as a promising alternative to autoregressive models, enabling parallel decoding fo

MoRe: Modular Representations for Principled Continual Representation Learning on Squantial Data

SafetyDGX agent

arXiv:2605.14364v1 Announce Type: new Abstract: Continual learning requires models to adapt to new data while preserving previously acquired knowledge. At its core, this challenge can be viewed as pri

Multi-Dimensional Model Integrity and Responsibility Assessment Index and Scoring Framework

SafetyDGX agent

arXiv:2605.14550v1 Announce Type: new Abstract: Artificial intelligence in high-stakes tabular domains cannot be evaluated by predictive performance alone, yet current practice still assesses explaina

Multi-scale Coarse-to-fine Modeling for Test-time Human Motion Control

SafetyDGX agent

arXiv:2605.14935v1 Announce Type: new Abstract: We present MSCoT, a multi-scale, coarse-to-fine model for test-time human motion synthesis and control. Unlike recent approaches that rely on multiple i

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning

SafetyDGX agent

arXiv:2512.07461v3 Announce Type: replace Abstract: We introduce Native Parallel Reasoner (NPR), a teacher-free framework that enables Large Language Models (LLMs) to self-evolve genuine parallel reas

NEST: Nested Event Stream Transformer for Sequences of Multisets

SafetyDGX agent

arXiv:2602.00520v3 Announce Type: replace Abstract: Event stream data often exhibit hierarchical structure in which multiple events co-occur, resulting in a sequence of multisets (i.e., bags of events

Not All Timesteps Matter Equally: Selective Alignment Knowledge Distillation for Spiking Neural Networks

SafetyDGX agent

arXiv:2605.14252v1 Announce Type: cross Abstract: Spiking neural networks (SNNs), which are brain-inspired and spike-driven, achieve high energy efficiency. However, a performance gap between SNNs and

On Strong Equivalence Notions in Logic Programming and Abstract Argumentation

SafetyDGX agent

arXiv:2605.14721v1 Announce Type: new Abstract: Strong equivalence between knowledge bases ensures the possibility of replacing one with the other without affecting reasoning outcomes, in any given co

On the Burden of Achieving Fairness in Conformal Prediction

SafetyDGX agent

arXiv:2605.14260v1 Announce Type: cross Abstract: Conformal prediction is often calibrated with a single pooled threshold, but this can hide cross-group heterogeneity in score distributions and distor

On the Unreasonable Effectiveness of Last-layer Retraining

SafetyDGX agent

arXiv:2512.01766v2 Announce Type: replace Abstract: Last-layer retraining (LLR) methods -- wherein the last layer of a neural network is reinitialized and retrained on a held-out set following ERM tra

OneSearch-V2: The Latent Reasoning Enhanced Self-distillation Generative Search Framework

SafetyDGX agent

arXiv:2603.24422v2 Announce Type: replace-cross Abstract: Generative Retrieval (GR) has emerged as a promising paradigm for modern search systems. Compared to multi-stage cascaded architecture, it off

Performance-Driven Policy Optimization for Speculative Decoding with Adaptive Windowing

SafetyDGX agent

arXiv:2605.14978v1 Announce Type: new Abstract: Speculative decoding accelerates LLM inference by having a lightweight draft model propose speculative windows of candidate tokens for parallel verifica

Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music

SafetyDGX agent

arXiv:2605.14765v1 Announce Type: cross Abstract: Persian music, with its unique tonalities, modal systems (Dastgah), and rhythmic structures, presents significant challenges for music generation mode

Phylogenetic Tree Inference with Tropical Axial Attention

SafetyDGX agent

arXiv:2605.13894v1 Announce Type: cross Abstract: In this work, we introduce a Tropical Axial Attention neural reasoning architecture that replaces vanilla softmax dot-product attention with max-plus

Policy Optimization in Hybrid Discrete-Continuous Action Spaces via Mixed Gradients

SafetyDGX agent

arXiv:2605.14297v1 Announce Type: cross Abstract: We study reinforcement learning in hybrid discrete-continuous action spaces, such as settings where the discrete component selects a regime (or index)

Pro-DG: Procedural Diffusion Guidance for Architectural Facade Generation

SafetyDGX agent

arXiv:2504.01571v2 Announce Type: replace-cross Abstract: We use hierarchical procedural rules for the generation of control maps within the stable diffusion framework to produce photo-realistic archi

Probabilistic Verification of Recurrent Neural Networks for Single and Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2605.14758v1 Announce Type: new Abstract: History-dependent policies induced by recurrent neural networks (RNNs) rely on latent hidden state dynamics, making verification in partially observable

Progent: Securing AI Agents with Privilege Control

SafetyDGX agent

arXiv:2504.11703v3 Announce Type: replace-cross Abstract: AI agents interact with external environments through tool calls, exposing them to attacks like indirect prompt injection that can trigger una

Prompting Policies for Multi-step Reasoning and Tool-Use in Black-box LLMs with Iterative Distillation of Experience

SafetyDGX agent

arXiv:2605.14443v1 Announce Type: new Abstract: The shift toward interacting with frozen, 'black-box' Large Language Models (LLMs) has transformed prompt engineering from a heuristic exercise into a c

Proximal Action Replacement for Behavior Cloning Actor-Critic in Offline Reinforcement Learning

SafetyDGX agent

arXiv:2602.07441v2 Announce Type: replace-cross Abstract: Offline reinforcement learning (RL), which optimizes policies using a previously collected static dataset, is an important branch of RL. A pop

Proxy Compression for Language Modeling

SafetyDGX agent

arXiv:2602.04289v2 Announce Type: replace Abstract: Modern language models are trained almost exclusively on token sequences produced by a fixed tokenizer, an external lossless compressor often over U

Quantitative Video World Model Evaluation for Geometric-Consistency

SafetyDGX agent

arXiv:2605.15185v1 Announce Type: cross Abstract: Generative video models are increasingly studied as implicit world models, yet evaluating whether they produce physically plausible 3D structure and m

R2PS: Worst-Case Robust Real-Time Pursuit Strategies under Partial Observability

SafetyDGX agent

arXiv:2511.17367v2 Announce Type: replace Abstract: Computing worst-case robust strategies in pursuit-evasion games (PEGs) is time-consuming, especially when real-world factors like partial observabil

R2R2: Robust Representation for Intensive Experience Reuse via Redundancy Reduction in Self-Predictive Learning

SafetyDGX agent

arXiv:2605.14026v1 Announce Type: cross Abstract: For reinforcement learning in data-scarce domains like real-world robotics, intensive data reuse enhances efficiency but induces overfitting. While pr

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling

SafetyDGX agent

arXiv:2510.20206v2 Announce Type: replace Abstract: Prompt design plays a crucial role in text-to-video (T2V) generation, yet user-provided prompts are often short, unstructured, and misaligned with t

RAVEN: Real-time Autoregressive Video Extrapolation with Consistency-model GRPO

SafetyDGX agent

arXiv:2605.15190v1 Announce Type: new Abstract: Causal autoregressive video diffusion models support real-time streaming generation by extrapolating future chunks from previously generated content. Di

Reasoning Model Is Superior LLM-Judge, Yet Suffers from Biases

SafetyDGX agent

arXiv:2601.03630v2 Announce Type: replace Abstract: This paper presents the first systematic comparison investigating whether Large Reasoning Models (LRMs) are superior judges to non-reasoning LLMs. O

Reinforcement Learning for Diffusion LLMs with Entropy-Guided Step Selection and Stepwise Advantages

SafetyDGX agent

arXiv:2603.12554v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has been effective for post-training autoregressive (AR) language models, but extending these methods to diffusion

Reinforcement Learning with Semantic Rewards Enables Low-Resource Language Expansion without Alignment Tax

SafetyDGX agent

arXiv:2605.14366v1 Announce Type: new Abstract: Extending large language models (LLMs) to low-resource languages often incurs an 'alignment tax': improvements in the target language come at the cost o

Resolving Action Bottleneck: Agentic Reinforcement Learning Informed by Token-Level Energy

SafetyDGX agent

arXiv:2605.14558v1 Announce Type: cross Abstract: Agentic reinforcement learning trains large language models using multi-turn trajectories that interleave long reasoning traces with short environment

Rethinking Output Alignment For 1-bit Post-Training Quantization of Large Language Models

SafetyDGX agent

arXiv:2512.21651v2 Announce Type: replace Abstract: Large Language Models (LLMs) deliver strong performance across a wide range of NLP tasks, but their massive sizes hinder deployment on resource-cons

REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding

SafetyDGX agent

arXiv:2511.13026v3 Announce Type: replace Abstract: Self-reflection mechanisms that rely on purely text-based rethinking processes perform well in most multimodal tasks. However, when directly applied

ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization

SafetyDGX agent

arXiv:2605.14497v1 Announce Type: cross Abstract: Offline-to-online reinforcement learning harnesses the stability of offline pretraining and the flexibility of online fine-tuning. A key challenge lie

Road Maps as Free Geometric Priors: Weather-Invariant Drone Geo-Localization with GeoFuse

SafetyDGX agent

arXiv:2605.14925v1 Announce Type: new Abstract: Drone-view geo-localization aims to match a query drone image, often captured under adverse weather conditions (e.g., rain, snow, fog), against a galler

Second-Order Actor-Critic Methods for Discounted MDPs via Policy Hessian Decomposition

SafetyDGX agent

arXiv:2605.14982v1 Announce Type: cross Abstract: We address the discounted reward setting in reinforcement learning (RL). To mitigate the value approximation challenges in policy gradient methods, ac

Self-Distilled Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.15155v1 Announce Type: cross Abstract: Reinforcement learning (RL) has emerged as a central paradigm for post-training LLM agents, yet its trajectory-level reward signal provides only coars

Send the arXiv AI-generated slop, get a yearlong vacation from submissions

SafetyDGX agent

ArXiv will ban authors for one year if they submit papers containing obviously AI-generated content , with examples including hallucinated citations, placeholder text, or chatbot meta-comments left in

SimPersona: Learning Discrete Buyer Personas from Raw Clickstreams for Grounded E-Commerce Agents

SafetyDGX agent

arXiv:2605.14205v1 Announce Type: new Abstract: LLM-based web agents can navigate live storefronts, yet they often collapse to a single 'average buyer' policy, failing to capture the heterogeneous and

SkillFlow: Flow-Driven Recursive Skill Evolution for Agentic Orchestration

SafetyDGX agent

arXiv:2605.14089v1 Announce Type: new Abstract: In recent years, a variety of powerful LLM-based agentic systems have been applied to automate complex tasks through task orchestration. However, existi

Slot-MPC: Goal-Conditioned Model Predictive Control with Object-Centric Representations

SafetyDGX agent

arXiv:2605.14937v1 Announce Type: cross Abstract: Predictive world models enable agents to model scene dynamics and reason about the consequences of their actions. Inspired by human perception, object

Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings

SafetyDGX agent

arXiv:2605.14284v1 Announce Type: new Abstract: Comparative evaluation of multiple dynamic treatment policies is essential for healthcare and policy decisions, yet conventional longitudinal causal inf

SOCC-ICP: Semantics-Assisted Odometry based on Occupancy Grids and ICP

SafetyDGX agent

arXiv:2605.15074v1 Announce Type: new Abstract: Reliable pose estimation in previously unseen environments is a fundamental capability of autonomous systems. Existing LiDAR odometry methods typically

SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

SafetyDGX agent

arXiv:2605.15044v1 Announce Type: cross Abstract: As audio-first agents become increasingly common in physical AI, conversational robots, and screenless wearables, audio large language models (audio-L

SpectraFlow: Unifying Structural Pretraining and Frequency Adaptation for Medical Image Segmentation

SafetyDGX agent

arXiv:2605.14566v1 Announce Type: new Abstract: Medical image segmentation remains challenging in low-data regimes, where scarce annotations often yield poor generalization and ambiguous boundaries wi

strawmanning Gary Marcus should be an Olympic sport, there are so many entrants. Hinton wins gold, for apparently faking a quote and putting…

SafetyDGX agent

Gary Marcus criticizes Geoffrey Hinton for allegedly misrepresenting or fabricating a quote attributed to him in a debate about AI. The post sarcastically compares the frequency of strawmanning argume

SuperF: Neural Implicit Fields for Multi-Image Super-Resolution

SafetyDGX agent

arXiv:2512.09115v2 Announce Type: replace Abstract: High-resolution imagery is often hindered by limitations in sensor technology, atmospheric conditions, and costs. Such challenges occur in satellite

Temporal Fair Division in Multi-Agent Systems: From Precise Alternation Metrics to Scalable Coordination Proxies

SafetyDGX agent

arXiv:2605.14879v1 Announce Type: cross Abstract: A plethora real-world environments require agents to compete repeatedly for the same limited resource, calling for a temporal notion of fairness judge

TERMS-Bench: Diagnosing LLM Negotiation Agents Beyond Deal Rate

SafetyDGX agent

arXiv:2605.13909v1 Announce Type: cross Abstract: Negotiation is a central mechanism of economic exchange, shaping markets, procurement, labor agreements, and resource allocation. It is also a canonic

The Great Pretender: A Stochasticity Problem in LLM Jailbreak

SafetyDGX agent

arXiv:2605.14418v1 Announce Type: cross Abstract: 'Oh-Oh, yes, I'm the great pretender. Pretending that I'm doing well. My need is such, I pretend too much...' summarizes the state in the area of jail

The Pitfalls of KV Cache Compression

SafetyDGX agent

arXiv:2510.00231v2 Announce Type: replace-cross Abstract: KV cache compression promises increased throughput and efficiency with negligible loss in performance. While the gains in throughput are indis

Towards Continuous Sign Language Conversation from Isolated Signs

SafetyDGX agent

arXiv:2605.14705v1 Announce Type: new Abstract: Sign language is the primary language for many Deaf and Hard-of-Hearing (DHH) signers, yet most conversational AI systems still mediate interaction thro

Training-Free Generative Sampling via Moment-Matched Score Smoothing

SafetyDGX agent

arXiv:2605.14276v1 Announce Type: cross Abstract: Diffusion models generate samples by denoising along the score of a perturbed target distribution. In practice, one trains a neural diffusion model, w

UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars

SafetyDGX agent

arXiv:2605.14731v1 Announce Type: cross Abstract: Speech-driven gestures and facial animations are fundamental to expressive digital avatars in games, virtual production, and interactive media. Howeve

Unbiased and Second-Order-Free Training for High-Dimensional PDEs

SafetyDGX agent

arXiv:2605.14643v1 Announce Type: new Abstract: Deep learning methods based on backward stochastic differential equations (BSDEs) have emerged as competitive alternatives to physics-informed neural ne

UniTriGen: Unified Triplet Generation of Aligned Visible-Infrared-Label for Few-Shot RGB-T Semantic Segmentation

SafetyDGX agent

arXiv:2605.14626v1 Announce Type: new Abstract: RGB-T semantic segmentation requires strictly aligned VIS-IR-Label triplets; however, such aligned triplet data are often scarce in real-world scenarios

← Previous
1…166167168169170…242
Next →