AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
13 May 2026

STRIDE: Training-Free Diversity Guidance via PCA-Directed Feature Perturbation in Single-Step Diffusion Models

SafetyDGX agent

arXiv:2605.11494v1 Announce Type: new Abstract: Distilled one-step (T=1) or few-step (Tleq4) diffusion models enable real-time image generation but often exhibit reduced sample diversity compared to t

SURGE: Surrogate Gradient Adaptation in Binary Neural Networks

SafetyDGX agent

arXiv:2605.10989v1 Announce Type: new Abstract: The training of Binary Neural Networks (BNNs) is fundamentally based on gradient approximation for non-differentiable binarization operations (e.g., sig

Tacmap: Bridging the Tactile Sim-to-Real Gap via Geometry-Consistent Penetration Depth Map

SafetyDGX agent

arXiv:2602.21625v2 Announce Type: replace Abstract: Vision-Based Tactile Sensors (VBTS) are essential for achieving dexterous robotic manipulation, yet the tactile sim-to-real gap remains a fundamenta


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Taming Extreme Tokens: Covariance-Aware GRPO with Gaussian-Kernel Advantage Reweighting

SafetyDGX agent

arXiv:2605.11538v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) has emerged as a promising approach for improving the reasoning capabilities of large language models. However

TAR: Text Semantic Assisted Cross-modal Image Registration Framework for Optical and SAR Images

SafetyDGX agent

arXiv:2605.12064v1 Announce Type: new Abstract: Existing deep learning-based methods can capture shared features from optical and synthetic aperture radar (SAR) images for spatial alignment. However,

Test-Time Personalization: A Diagnostic Framework and Probabilistic Fix for Scaling Failures

SafetyDGX agent

arXiv:2605.10991v1 Announce Type: new Abstract: Existing approaches to LLM personalization focus on constructing better personalized models or inputs, while treating inference as a single-shot process

The DAWN of World-Action Interactive Models

SafetyDGX agent

arXiv:2605.11550v1 Announce Type: new Abstract: A plausible scene evolution depends on the maneuver being considered, while a good maneuver depends on how the scene may evolve. Existing World Action M

The Evaluation Differential: When Frontier AI Models Recognise They Are Being Tested

SafetyDGX agent

arXiv:2605.11496v1 Announce Type: cross Abstract: Recent published evidence from frontier laboratories shows that contemporary AI models can recognise evaluation contexts, latently represent them, and

The new era of SaMD: Why cloud infrastructure is the foundation for digital health in 2026

SafetyDGX agent

In the healthcare and life sciences industries, speed saves lives, but meeting regulatory requirements and other administrative burdens often pumps the brakes for manufacturers of software as a medica

The tractability landscape of diffusion alignment: regularization, rewards, and computational primitives

SafetyDGX agent

arXiv:2605.11361v1 Announce Type: new Abstract: Inference-time reward alignment asks how to turn a pre-trained diffusion model with base law p into a sampler that favors a reward r while remaining clo

there are many agent use cases locked behind 'what if' fears because after we give a tool to an agent, we're relying on the prompt to limit …

SafetyDGX agent

there are many agent use cases locked behind 'what if' fears because after we give a tool to an agent, we're relying on the prompt to limit behavior tools like @denieddotdev allow teams to manage capa

TMPO: Trajectory Matching Policy Optimization for Diverse and Efficient Diffusion Alignment

SafetyDGX agent

arXiv:2605.10983v1 Announce Type: cross Abstract: Reinforcement learning (RL) has shown extraordinary potential in aligning diffusion models to downstream tasks, yet most of them still suffer from sig

TMRL: Diffusion Timestep-Modulated Pretraining Enables Exploration for Efficient Policy Finetuning

SafetyDGX agent

arXiv:2605.12236v1 Announce Type: cross Abstract: Fine-tuning pre-trained robot policies with reinforcement learning (RL) often inherits the bottlenecks introduced by pre-training with behavioral clon

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching

SafetyDGX agent

arXiv:2605.12288v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is a widely used RL-free method for aligning language models from pairwise preferences, but it models preferences o

Towards Fine-Grained Code-Switch Speech Translation with Semantic Space Alignment

SafetyDGX agent

arXiv:2511.10670v2 Announce Type: replace Abstract: Code-switching (CS) speech translation (ST) aims to translate speech that alternates between multiple languages into a target language text, posing

Towards Order Fairness: Mitigating LLMs Order Sensitivity through Dual Group Advantage Optimization

SafetyDGX agent

arXiv:2605.11974v1 Announce Type: new Abstract: Large Language Models (LLMs) suffer from order bias, where their performance is affected by the arrangement order of input elements. This unfairness lim

Toxicity Detection Should Measure Contextual Harm, Not Text-Intrinsic Badness

SafetyDGX agent

arXiv:2503.16072v4 Announce Type: replace-cross Abstract: Toxicity detection has become core safety infrastructure for online moderation, dataset filtering, and deployed language-model systems. Yet mo

Training Transformers for KV Cache Compressibility

SafetyDGX agent

arXiv:2605.05971v2 Announce Type: replace Abstract: Long-context language modeling is increasingly constrained by the Key-Value (KV) cache, whose memory and decode-time access costs scale linearly wit

Trajectory First: A Curriculum for Discovering Diverse Policies

SafetyDGX agent

arXiv:2506.01568v3 Announce Type: replace Abstract: Being able to solve a task in diverse ways makes agents more robust to task variations and less prone to local optima. In this context, constrained

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

SafetyDGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

Transformer-Based Autonomous Driving Models and Deployment-Oriented Compression: A Survey

SafetyDGX agent

arXiv:2304.10891v2 Announce Type: replace-cross Abstract: Transformer-based models are becoming a central paradigm in autonomous driving because they can capture long-range spatial dependencies, multi

Trust Region Inverse Reinforcement Learning: Explicit Dual Ascent using Local Policy Updates

SafetyDGX agent

arXiv:2605.11020v1 Announce Type: new Abstract: Inverse reinforcement learning (IRL) is typically formulated as maximizing entropy subject to matching the distribution of expert trajectories. Classica

Trust the Batch, On- or Off-Policy: Adaptive Policy Optimization for RL Post-Training

SafetyDGX agent

arXiv:2605.12380v1 Announce Type: new Abstract: Reinforcement learning is structurally harder than supervised learning because the policy changes the data distribution it learns from. The resulting fr

UGround: Towards Unified Visual Grounding with Unrolled Transformers

SafetyDGX agent

arXiv:2510.03853v4 Announce Type: replace Abstract: We present UGround, a extbf{U}nified visual extbf{Ground}ing paradigm that dynamically selects intermediate layers across extbf{U}nrolled transforme

Understanding and Preventing Entropy Collapse in RLVR with On-Policy Entropy Flow Optimization

SafetyDGX agent

arXiv:2605.11491v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has become an effective paradigm for improving the reasoning ability of large language models. How

Understanding Sample Efficiency in Predictive Coding

SafetyDGX agent

arXiv:2605.11911v1 Announce Type: new Abstract: Predictive Coding (PC) is an influential account of cortical learning. Much of recent work has focused on comparing PC to Backpropagation (BP) to find w

Understanding the Performance Gap in Preference Learning: A Dichotomy of RLHF and DPO

SafetyDGX agent

arXiv:2505.19770v5 Announce Type: replace-cross Abstract: We present a fine-grained theoretical analysis of the performance gap between two-stage reinforcement learning from human feedback~(RLHF) and

UniFixer: A Universal Reference-Guided Fixer for Diffusion-Based View Synthesis

SafetyDGX agent

arXiv:2605.12169v1 Announce Type: new Abstract: With the recent surge of generative models, diffusion-based approaches have become mainstream for view synthesis tasks, either in an explicit depth-warp

VIP: Visual-guided Prompt Evolution for Efficient Dense Vision-Language Inference

SafetyDGX agent

arXiv:2605.12325v1 Announce Type: new Abstract: Pursuing training-free open-vocabulary semantic segmentation in an efficient and generalizable manner remains challenging due to the deep-seated spatial

VNDUQE: Information-Theoretic Novelty Detection using Deep Variational Information Bottleneck

SafetyDGX agent

arXiv:2605.11551v1 Announce Type: cross Abstract: Detecting out-of-distribution (OOD) samples is critical for safe deployment of neural networks in safety-critical applications. While maximum softmax

way ahead of its time:

SafetyDGX agent

way ahead of its time: Three questions for @sama that the public deserves to better understand: 👉 What is current value of your indirect stake in OpenAI? (Note that you told the senate that you had no

What-Where Transformer: A Slot-Centric Visual Backbone for Concurrent Representation and Localization

SafetyDGX agent

arXiv:2605.12021v1 Announce Type: new Abstract: Many image understanding tasks involve identifying what is present and where it appears. However, tasks that address where, such as object discovery, de

When Does ell_2-Boosting Overfit Benignly? High-Dimensional Risk Asymptotics and the ell_1 Implicit Bias

SafetyDGX agent

arXiv:2605.06314v2 Announce Type: replace Abstract: Benign overfitting is well-characterized in ell_2 geometries, but its behavior under the ell_1 implicit bias of greedy ensembles remains challenging

When Policy Entropy Constraint Fails: Preserving Diversity in Flow-based RLHF via Perceptual Entropy

SafetyDGX agent

arXiv:2605.12112v1 Announce Type: new Abstract: RLHF is widely used to align flow-matching text-to-image models with human preferences, but often leads to severe diversity collapse after fine-tuning.

When to Ask a Question: Understanding Communication Strategies in Generative AI Tools

SafetyDGX agent

arXiv:2605.11240v1 Announce Type: cross Abstract: Generative AI models differ from traditional machine learning tools in that they allow users to provide as much or as little information as they choos

World Action Models: The Next Frontier in Embodied AI

SafetyDGX agent

arXiv:2605.12090v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have achieved strong semantic generalization for embodied policy learning, yet they learn reactive observation-to-

ZeroIDIR: Zero-Reference Illumination Degradation Image Restoration with Perturbed Consistency Diffusion Models

SafetyDGX agent

arXiv:2605.11435v1 Announce Type: new Abstract: In this paper, we propose a zero-reference diffusion-based framework, named ZeroIDIR, for illumination degradation image restoration, which decouples th

12 May 2026

A Cross-Layered Multi-Drone Coordination for Medical Supply Delivery during Disaster Response Management

SafetyDGX agent

arXiv:2605.09342v1 Announce Type: cross Abstract: Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in

A Scalable Entity-Based Framework for Auditing Bias in LLMs

SafetyDGX agent

arXiv:2601.12374v2 Announce Type: replace-cross Abstract: Existing approaches to bias evaluation in large language models (LLMs) trade ecological validity for statistical control, relying either on ar

A Single Deep Preference-Conditioned Policy for Learning Pareto Coverage Sets

SafetyDGX agent

arXiv:2605.08946v1 Announce Type: new Abstract: Preference-conditioned multi-objective reinforcement learning aims to learn a single policy that captures trade-offs across preferences, but under nonli

A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models

SafetyDGX agent

arXiv:2605.08513v1 Announce Type: cross Abstract: Safety alignment in language models operates through two mechanistically distinct systems: refusal neurons that gate whether harmful knowledge is expr

A true exponential!

SafetyDGX agent

A true exponential! Oy. According to a new paper in The Lancet, the rate of made-up citations in biomedical papers has increased by more than 12x since 2023. https://www.thelancet.com/journals/lancet/

ActivationReasoning: Logical Reasoning in Latent Activation Spaces

SafetyDGX agent

arXiv:2510.18184v3 Announce Type: replace-cross Abstract: Large language models (LLMs) excel at generating fluent text, but their internal reasoning remains opaque and difficult to control. Sparse aut

Active Tabular Augmentation via Policy-Guided Diffusion Inpainting

SafetyDGX agent

arXiv:2605.10315v1 Announce Type: cross Abstract: Generative tabular augmentation is appealing in data-scarce domains, yet the prevailing focus on distributional fidelity does not reliably translate i

Adaptive Context Matters: Towards Provable Multi-Modality Guidance for Super-Resolution

SafetyDGX agent

arXiv:2605.10470v1 Announce Type: new Abstract: Super-resolution (SR) is a severely ill-posed problem with inherent ambiguity, as widely recognized in both empirical and theoretical studies. Although

Adaptive Data Harvesting for Efficient Neural Network Learning with Universal Constraints

SafetyDGX agent

arXiv:2605.09707v1 Announce Type: cross Abstract: Training neural networks to satisfy universal constraints over continuous domains poses unique challenges. Common examples include Lyapunov Neural Net

Adversarial Attacks Against MLLMs via Progressive Resolution Processing and Adaptive Feature Alignment

SafetyDGX agent

arXiv:2605.09902v1 Announce Type: new Abstract: Adversarial perturbations can mislead Multimodal Large Language Models (MLLMs) recognize a benign image as a specific target object, posing serious risk

Agent-Omit: Adaptive Context Omission for Efficient LLM Agents

SafetyDGX agent

arXiv:2602.04284v2 Announce Type: replace Abstract: Managing agent context (e.g., thought and observation) during multi-turn agent-environment interactions is an emerging strategy to improve agent eff

Agent-Sentry: Bounding LLM Agents via Execution Provenance

SafetyDGX agent

arXiv:2603.22868v2 Announce Type: replace-cross Abstract: Agentic computing systems, while immensely capable, raise serious security, privacy, and safety concerns. A key issue is that the full set of

AgentReview: Exploring Peer Review Dynamics with LLM Agents

SafetyDGX agent

arXiv:2406.12708v3 Announce Type: replace Abstract: Peer review is fundamental to the integrity and advancement of scientific publication. Traditional methods of peer review analyses often rely on exp

AI-Care: A Conversational Agentic System for Task Coordination in Alzheimer's Disease Care

SafetyDGX agent

arXiv:2605.08480v1 Announce Type: new Abstract: Individuals with Alzheimer's disease (AD) and Alzheimer's disease-related dementia (ADRD) experience memory and thinking changes that impact their abili

AIPO: : Learning to Reason from Active Interaction

SafetyDGX agent

arXiv:2605.08401v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have demonstrated remarkable reasoning capabilities, largely stimulated by Reinforcement Learning with

ALAM: Algebraically Consistent Latent Transitions for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.10819v1 Announce Type: cross Abstract: Vision-language-action (VLA) models remain constrained by the scarcity of action-labeled robot data, whereas action-free videos provide abundant evide

Align and Shine: Building High-Quality Sentence-Aligned Corpora for Multilingual Text Simplification

SafetyDGX agent

arXiv:2605.09476v1 Announce Type: cross Abstract: Text simplification plays a crucial role in improving the accessibility and comprehensibility of written information for diverse audiences, including

Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis

SafetyDGX agent

arXiv:2605.10415v1 Announce Type: new Abstract: Large language models for subjectivity analysis are typically trained with aggregated labels, which compress variations in human judgment into a single

Aligning Validation with Deployment: Target-Weighted Cross-Validation for Spatial Prediction

SafetyDGX agent

arXiv:2603.29981v2 Announce Type: replace Abstract: Reliable estimation of predictive performance is essential for spatial environmental modeling, where machine-learning models are used to generate ma

Alignment as Jurisprudence

SafetyDGX agent

arXiv:2605.08416v1 Announce Type: new Abstract: Jurisprudence, the study of how judges should properly decide cases, and alignment, the science of getting AI models to conform to human values, share a

Alignment-Sensitive Minimax Rates for Spectral Algorithms with Learned Kernels

SafetyDGX agent

arXiv:2509.20294v4 Announce Type: replace Abstract: We study spectral algorithms in the setting where kernels are learned from data. We introduce the effective span dimension (ESD), an alignment-sensi

An Empirical Analysis of Calibration and Selective Prediction in Multimodal Clinical Condition Classification

SafetyDGX agent

arXiv:2603.02719v2 Announce Type: replace Abstract: As artificial intelligence systems move toward clinical deployment, ensuring reliable prediction behavior is fundamental for safety-critical decisio

Anatomical Landmark-Guided Deep Reinforcement Learning for Autonomous Gastric Navigation

SafetyDGX agent

arXiv:2605.08269v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables painless visualization of the gastrointestinal tract, but its diagnostic potential is limited by incomplete muc

← Previous
1…146147148149150…212
Next →