AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
21 May 2026

Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization

SafetyDGX agent

arXiv:2510.21583v2 Announce Type: replace Abstract: Recent Progress in post-training flow matching for text-to-image (T2I) generation with Group Relative Policy Optimization (GRPO) has demonstrated st

Proximal State Nudging: Reducing Skill Atrophy from AI Assistance

SafetyDGX agent

arXiv:2605.20355v1 Announce Type: cross Abstract: Skill atrophy, the gradual decline of human capability under AI assistance, poses a safety risk in shared-control of semi-autonomous systems, where op

Q-SpiRL: Quantum Spiking Reinforcement Learning for Adaptive Robot Navigation

SafetyDGX agent

arXiv:2605.20801v1 Announce Type: new Abstract: Adaptive robot navigation in dynamic environments requires policies that can reach the target reliably while producing efficient and stable trajectories


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Quadratic Characterizations for Reachability Analysis of Neural Networks

SafetyDGX agent

arXiv:2605.20482v1 Announce Type: new Abstract: Quadratic constraints (QCs) are widely used to characterize nonlinearities and uncertainties, but generic analytical characterizations can be conservati

Quantum End-to-End Learning for Contextual Combinatorial Optimization

SafetyDGX agent

arXiv:2605.20222v1 Announce Type: cross Abstract: Contextual combinatorial optimization (CCO) plays a critical role in decision-making under uncertainty, yet remains a significant challenge. We presen

REFLECTOR: Internalizing Step-wise Reflection against Indirect Jailbreak

SafetyDGX agent

arXiv:2605.20654v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable capabilities, they remain susceptible to sophisticated, multi-step jailbreak attacks that circ

Regulating Anatomy-Aware Rewards via Trajectory-Integral Feedback for Volumetric Computed Tomography Analysis

SafetyDGX agent

arXiv:2605.20277v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have rapidly advanced as general-purpose multimodal assistants, yet their deployment in 3D Computed Tomography (CT

Reinforcement Learning for Risk Adaptation via Differentiable CVaR Barrier Functions

SafetyDGX agent

arXiv:2605.21257v1 Announce Type: new Abstract: Planning through crowded environments under uncertain obstacle motions remains difficult, as stochastic interactions often induce overly conservative be

Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces

SafetyDGX agent

arXiv:2509.22963v3 Announce Type: replace Abstract: Reinforcement learning (RL) struggles to scale to large, combinatorial action spaces common in many real-world problems. This paper introduces a nov

rePIRL: Learn PRM with Inverse RL for LLM Reasoning

SafetyDGX agent

arXiv:2602.07832v2 Announce Type: replace Abstract: Process rewards have been widely used in deep reinforcement learning to improve training efficiency, reduce variance, and prevent reward hacking. In

Rethinking Cross-Layer Information Routing in Diffusion Transformers

SafetyDGX agent

arXiv:2605.20708v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have become a de facto backbone of modern visual generation, and nearly every major axis of their design -- tokenization,

Robust Recommendation from Noisy Implicit Feedback: A GMM-Weighted Bayes-label Transition Matrix Framework

SafetyDGX agent

arXiv:2605.20721v1 Announce Type: new Abstract: Learning from implicit feedback in recommender systems is fundamentally challenged by pervasive label noise. While conventional denoising approaches oft

SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection

SafetyDGX agent

arXiv:2605.21186v1 Announce Type: new Abstract: Interpretability in object detection provides crucial confidence support for clinical auxiliary diagnosis. However, in tiny bacteria detection, traditio

SCRIBE: Diagnostic Evaluation and Rich Transcription Models for Indic ASR

SafetyDGX agent

arXiv:2605.20712v1 Announce Type: new Abstract: Automatic speech recognition replaces typing only when correction costs less than manual entry, a threshold determined by error types, not counts: fixin

Secure, Verifiable, and Scalable Multi-Client Data Sharing via Consensus-Based Privacy-Preserving Data Distribution

SafetyDGX agent

arXiv:2601.00418v2 Announce Type: replace-cross Abstract: We propose the Consensus-Based Privacy-Preserving Data Distribution (CPPDD) framework, a lightweight and post-setup autonomous protocol for se

Self-Refining Video Sampling

SafetyDGX agent

arXiv:2601.18577v2 Announce Type: replace Abstract: Modern video generators still struggle with complex physical dynamics, often falling short of physical realism. Existing approaches address this usi

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

SafetyDGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

Shipping features to production just got easier with new feature flags in AppLifecycle Manager

SafetyDGX agent

Many development teams are familiar with the hesitation that comes right before pushing a new feature live. As AI helps developers write code faster, the gap between rapid code generation and safe pro

SMA-DP: Spectral Memory-Aware Differential Privacy for Deep Learning

SafetyDGX agent

arXiv:2605.20450v1 Announce Type: new Abstract: Differentially private stochastic gradient descent (DP-SGD) enables private deep learning through per-example clipping and calibrated Gaussian noise, bu

Spacetime Optimal-Transport Attention for Visuo-Haptic Imitation Learning of Contact-Rich Manipulation

SafetyDGX agent

arXiv:2605.20433v1 Announce Type: new Abstract: Contact-rich manipulation tasks such as tight-clearance insertion, connector mating, polishing, and surface-conforming wiping remain difficult for data-

SpaceX claims a 28.5 TRILLION market let us look at Musk's pitch deck to banks for the twitter takeover in 2022, he said he'd create a 10 …

SafetyDGX agent

SpaceX claims a 28.5 TRILLION market let us look at Musk's pitch deck to banks for the twitter takeover in 2022, he said he'd create a 10 billion dollar subs biz by 2028 today, ad revs down 75% & subs

Spatial Gram Alignment for Ultra-High-Resolution Image Synthesis

SafetyDGX agent

arXiv:2605.20808v1 Announce Type: new Abstract: Modern ultra-high-resolution image synthesis relies heavily on the robust generative capacity of large-scale pre-trained Latent Diffusion Models (LDMs).

Spectral Souping: A Unified Framework for Online Preference Alignment

SafetyDGX agent

arXiv:2605.20408v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) effectively aligns Large Language Models (LLMs) with aggregate human preferences but often fails to ad

Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables

SafetyDGX agent

arXiv:2605.20478v1 Announce Type: new Abstract: LLM-curated tables can appear source-grounded while containing unsupported rows: the curator may recall entries from parametric memory and retroactively

Statistical Guarantees in the Search for Less Discriminatory Algorithms

SafetyDGX agent

arXiv:2512.23943v2 Announce Type: replace-cross Abstract: U.S. discrimination law can impose liability on firms that fail to adopt a less discriminatory alternative (LDA): a decision policy that achie

STEAM: A Training-Free Congestion-Aware Enhancement Framework for Decentralized Multi-Agent Path Finding

SafetyDGX agent

arXiv:2605.20929v1 Announce Type: new Abstract: We propose STEAM (Spatial, Temporal, and Emergent congestion Awareness for MAPF), a training-free test-time enhancement framework for learning-based dec

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2605.21261v1 Announce Type: new Abstract: Training-free zero-shot composed image retrieval models are recently gaining increasing research interest due to their generalizability and flexibility

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning

SafetyDGX agent

arXiv:2506.21039v3 Announce Type: replace Abstract: Long-horizon goal-conditioned tasks pose fundamental challenges for reinforcement learning (RL), particularly when goals are distant and rewards are

Subword boundaries are the second meaningful effect. Adding end-of-subword markers as input embeddings produces a large gain throughout trai…

SafetyDGX agent

Subword boundaries are the second meaningful effect. Adding end-of-subword markers as input embeddings produces a large gain throughout training (H3): end-boundaries leak future bytes (whitespace alwa

SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework

SafetyDGX agent

arXiv:2605.20373v1 Announce Type: cross Abstract: Building humanoid robots capable of generalizable whole-body loco-manipulation in the real world remains a fundamental challenge. Existing methods eit

Supervised Latent Restructuring for Small-Data Quantum Learning in Plant Phenomics

SafetyDGX agent

arXiv:2605.20413v1 Announce Type: new Abstract: High-dimensional biological data often exhibit a severe mismatch between feature dimensionality and sample size, making reliable classification difficul

SURF: Steering the Scalarization Weight to Uniformly Traverse the Pareto Front

SafetyDGX agent

arXiv:2605.20619v1 Announce Type: new Abstract: Scalarization is widely used in multi-objective optimization owing to its simplicity and scalability. In many applications, the goal is to generate solu

SurgOnAir: Hierarchy-Aware Real-Time Surgical Video Commentary

SafetyDGX agent

arXiv:2605.21132v1 Announce Type: new Abstract: Understanding surgical workflow in real time is fundamental for intelligent surgical embodiment, where AI systems continuously perceive and respond as s

SynCB: A Synergy Concept-Based Model with Dynamic Routing Between Concepts and Complementary Neural Branches

SafetyDGX agent

arXiv:2605.20908v1 Announce Type: new Abstract: Concept-based (CB) models provide interpretability and support test-time human intervention, while standard neural networks (NN) offer strong task perfo

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

SafetyDGX agent

arXiv:2605.20356v1 Announce Type: new Abstract: Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-base

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

SafetyDGX agent

arXiv:2506.08277v3 Announce Type: replace-cross Abstract: Recent voxel-wise multimodal brain encoding studies have shown that multimodal large language models (MLLMs) exhibit a higher degree of brain

Tech researchers are suing the Trump administration over the future of online safety

SafetyDGX agent

Since its earliest days back in office, the Trump administration has been going after researchers who study and try to counter hate speech, harassment, propaganda, and disinformation online. Now, some

TelePhysics: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction

SafetyDGX agent

arXiv:2605.20290v1 Announce Type: cross Abstract: Recent generative video models achieve impressive visual quality but remain constrained by limited physical consistency and controllability. Existing

The Download: online safety’s future and climate tech’s big pivot

SafetyDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Tech researchers are suing the Trump administration over the f

The Economics of AI Inference: Inflation Dynamics, Welfare Costs, and Optimal Monetary Policy under the Inference-Cost Phillips Curve

SafetyDGX agent

arXiv:2605.20281v1 Announce Type: cross Abstract: We develop a unified microeconomic and monetary theory of artificial intelligence inference costs and their pass-through to inflation, welfare, and op

The Illusion of Intervention: Your LLM-Simulated Experiment is an Observational Study

SafetyDGX agent

arXiv:2605.20767v1 Announce Type: new Abstract: Large language models (LLMs) show potential as simulators of human behavior, offering a scalable way to study responses to interventions. However, becau

Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting

SafetyDGX agent

arXiv:2506.17631v4 Announce Type: replace Abstract: Time series forecasting aims to model temporal dependencies among variables for future state inference, holding significant importance and widesprea

Time-To-Reach Separation and Safety Filtering for Safe, Fair, and Efficient Multi-Agent Coordination

SafetyDGX agent

arXiv:2605.20625v1 Announce Type: cross Abstract: Advanced Air Mobility (AAM) operations are expected to significantly increase aerial traffic in urban airspace, requiring autonomous traffic managemen

Towards Context-Invariant Safety Alignment for Large Language Models

SafetyDGX agent

arXiv:2605.20994v1 Announce Type: new Abstract: Preference-based post-training aligns LLMs with human intent, yet safety behavior often remains brittle. A model may refuse a harmful request in a stand

Trusted Weights, Treacherous Optimizations? Optimization-Triggered Backdoor Attacks on LLMs

SafetyDGX agent

arXiv:2605.20641v1 Announce Type: cross Abstract: Inference optimization is a vital technique for deploying LLMs at scale. Compilation is the most widely adopted optimization technique for LLMs. While

Tutor-Student Reinforcement Learning: A Dynamic Curriculum for Robust Deepfake Detection

SafetyDGX agent

arXiv:2603.24139v2 Announce Type: replace Abstract: Standard supervised training for deepfake detection treats all samples with uniform importance, which can be suboptimal for learning robust and gene

Uh-oh, the International Space Station is leaking again

SafetyDGX agent

The International Space Station has an ongoing air leak in its Russian PrK module that has been known since 2019 , and this leak has delayed the launch of Axiom Space's private astronaut mission . NAS

Uncertainty-Calibrated Explainable Artificial Intelligence for Fetal Ultrasound Plane Classification: A Systematic Review

SafetyDGX agent

arXiv:2601.00990v2 Announce Type: replace-cross Abstract: Fetal ultrasound is the cornerstone of antenatal care, and accurate recognition of a small set of standard anatomical planes underpins biometr

Velocityformer: Broken-Symmetry-Matched Equivariant Graph Transformers for Cosmological Velocity Reconstruction

SafetyDGX agent

arXiv:2605.21483v1 Announce Type: cross Abstract: Precise measurement of the kinematic Sunyaev-Zel'dovich (kSZ) effect - a probe of the large-scale distribution of baryonic matter, a key observable fo

Verifiable Error Bounds for Physics-Informed Neural Network Solutions of Lyapunov and Hamilton-Jacobi-Bellman Equations

SafetyDGX agent

arXiv:2603.19545v2 Announce Type: replace-cross Abstract: Many core problems in nonlinear systems analysis and control can be recast as solving partial differential equations (PDEs) such as Lyapunov a

VLANeXt: Recipes for Building Strong VLA Models

SafetyDGX agent

arXiv:2602.18532v2 Announce Type: replace Abstract: Following the rise of large foundation models, Vision-Language-Action models (VLAs) emerged, leveraging strong visual and language understanding fro

Waymo suspends freeway rides and pauses its Atlanta operations, as it updates software to improve performance around construction zones and flooded roadways (Reuters)

SafetyDGX agent

Reuters: Waymo suspends freeway rides and pauses its Atlanta operations, as it updates software to improve performance around construction zones and flooded roadways — Alphabet's (GOOGL.O) Waymo said

What Semantics Survive the Connector? Diagnosing VLM-to-DiT Alignment in Video Editing

SafetyDGX agent

arXiv:2605.20795v1 Announce Type: new Abstract: Flow matching based video generative models have been increasingly relying on prepended Vision-Language Models (VLMs) to handle complex, instruction-bas

When AI Gets it Wrong: Reliability and Risk in AI-Assisted Medication Decision Systems

SafetyDGX agent

arXiv:2604.01449v3 Announce Type: replace-cross Abstract: Artificial intelligence (AI) systems are increasingly integrated into healthcare and pharmacy workflows, supporting tasks such as medication r

When does circular financing stop? Similar to a Ponzi Scheme, when cash outflows become greater than inflows. In this last quarter, $NVDA ca…

SafetyDGX agent

When does circular financing stop? Similar to a Ponzi Scheme, when cash outflows become greater than inflows. In this last quarter, NVDA cash grew by only ~600, and now ~95% of its operating cash flow

Why Aggregate Accuracy is Inadequate for Evaluating Fairness in Law Enforcement Facial Recognition Systems

SafetyDGX agent

arXiv:2603.28675v2 Announce Type: replace Abstract: Facial recognition systems are increasingly deployed in law enforcement and security contexts, where algorithmic decisions can carry significant soc

Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k Experts

SafetyDGX agent

arXiv:2504.12988v5 Announce Type: replace Abstract: Existing Learning-to-Defer (L2D) frameworks are limited to single-expert deferral, forcing each query to rely on only one expert and preventing the

Wonder what fraction of those liking this realized it was a joke 🤣

SafetyDGX agent

Wonder what fraction of those liking this realized it was a joke 🤣 There's a lot of excitement about OpenAI's resolution of the unit distance conjecture. In this thread I will discuss the striking app

You can be a cheerleader, or you can be a scientist. Can’t be both. cc: @scaling01

SafetyDGX agent

Gary Marcus argues that pursuing cheerleading and science as careers are mutually exclusive paths, suggesting fundamental incompatibility between these pursuits. The post likely critiques the idea of

Yowza! Similar in some ways to the Stanford paper recently on LLMs hallucinating responses to images they never saw.

SafetyDGX agent

Yowza! Similar in some ways to the Stanford paper recently on LLMs hallucinating responses to images they never saw. This was an amazing and incredibly damning experiment using Microsoft Copilot, by @

← Previous
1…127128129130131…214
Next →