AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
21 May 2026

rePIRL: Learn PRM with Inverse RL for LLM Reasoning

SafetyDGX agent

arXiv:2602.07832v2 Announce Type: replace Abstract: Process rewards have been widely used in deep reinforcement learning to improve training efficiency, reduce variance, and prevent reward hacking. In

Rethinking Cross-Layer Information Routing in Diffusion Transformers

SafetyDGX agent

arXiv:2605.20708v1 Announce Type: new Abstract: Diffusion Transformers (DiTs) have become a de facto backbone of modern visual generation, and nearly every major axis of their design -- tokenization,

Robust Recommendation from Noisy Implicit Feedback: A GMM-Weighted Bayes-label Transition Matrix Framework

SafetyDGX agent

arXiv:2605.20721v1 Announce Type: new Abstract: Learning from implicit feedback in recommender systems is fundamentally challenged by pervasive label noise. While conventional denoising approaches oft

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection

SafetyDGX agent

arXiv:2605.21186v1 Announce Type: new Abstract: Interpretability in object detection provides crucial confidence support for clinical auxiliary diagnosis. However, in tiny bacteria detection, traditio

SCRIBE: Diagnostic Evaluation and Rich Transcription Models for Indic ASR

SafetyDGX agent

arXiv:2605.20712v1 Announce Type: new Abstract: Automatic speech recognition replaces typing only when correction costs less than manual entry, a threshold determined by error types, not counts: fixin

Secure, Verifiable, and Scalable Multi-Client Data Sharing via Consensus-Based Privacy-Preserving Data Distribution

SafetyDGX agent

arXiv:2601.00418v2 Announce Type: replace-cross Abstract: We propose the Consensus-Based Privacy-Preserving Data Distribution (CPPDD) framework, a lightweight and post-setup autonomous protocol for se

Self-Refining Video Sampling

SafetyDGX agent

arXiv:2601.18577v2 Announce Type: replace Abstract: Modern video generators still struggle with complex physical dynamics, often falling short of physical realism. Existing approaches address this usi

Shiny Stories, Hidden Struggles: Investigating the Representation of Disability Through the Lens of LLMs

SafetyDGX agent

arXiv:2605.20191v1 Announce Type: new Abstract: Modern Large Language Models (LLMs) have recently attracted much attention for their ability to simulate human behavior and generate text that reflects

SMA-DP: Spectral Memory-Aware Differential Privacy for Deep Learning

SafetyDGX agent

arXiv:2605.20450v1 Announce Type: new Abstract: Differentially private stochastic gradient descent (DP-SGD) enables private deep learning through per-example clipping and calibrated Gaussian noise, bu

SpaceX claims a 28.5 TRILLION market let us look at Musk's pitch deck to banks for the twitter takeover in 2022, he said he'd create a 10 …

SafetyDGX agent

SpaceX claims a 28.5 TRILLION market let us look at Musk's pitch deck to banks for the twitter takeover in 2022, he said he'd create a 10 billion dollar subs biz by 2028 today, ad revs down 75% & subs

Spatial Gram Alignment for Ultra-High-Resolution Image Synthesis

SafetyDGX agent

arXiv:2605.20808v1 Announce Type: new Abstract: Modern ultra-high-resolution image synthesis relies heavily on the robust generative capacity of large-scale pre-trained Latent Diffusion Models (LDMs).

Spectral Souping: A Unified Framework for Online Preference Alignment

SafetyDGX agent

arXiv:2605.20408v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) effectively aligns Large Language Models (LLMs) with aggregate human preferences but often fails to ad

Stage-Audit: Auditable Source-Frontier Discovery for Cross-Wiki Tables

SafetyDGX agent

arXiv:2605.20478v1 Announce Type: new Abstract: LLM-curated tables can appear source-grounded while containing unsupported rows: the curator may recall entries from parametric memory and retroactively

Statistical Guarantees in the Search for Less Discriminatory Algorithms

SafetyDGX agent

arXiv:2512.23943v2 Announce Type: replace-cross Abstract: U.S. discrimination law can impose liability on firms that fail to adopt a less discriminatory alternative (LDA): a decision policy that achie

STEAM: A Training-Free Congestion-Aware Enhancement Framework for Decentralized Multi-Agent Path Finding

SafetyDGX agent

arXiv:2605.20929v1 Announce Type: new Abstract: We propose STEAM (Spatial, Temporal, and Emergent congestion Awareness for MAPF), a training-free test-time enhancement framework for learning-based dec

STiTch: Semantic Transition and Transportation in Collaboration for Training-Free Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2605.21261v1 Announce Type: new Abstract: Training-free zero-shot composed image retrieval models are recently gaining increasing research interest due to their generalizability and flexibility

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning

SafetyDGX agent

arXiv:2506.21039v3 Announce Type: replace Abstract: Long-horizon goal-conditioned tasks pose fundamental challenges for reinforcement learning (RL), particularly when goals are distant and rewards are

Subword boundaries are the second meaningful effect. Adding end-of-subword markers as input embeddings produces a large gain throughout trai…

SafetyDGX agent

Subword boundaries are the second meaningful effect. Adding end-of-subword markers as input embeddings produces a large gain throughout training (H3): end-boundaries leak future bytes (whitespace alwa

SUGAR: A Scalable Human-Video-Driven Generalizable Humanoid Loco-Manipulation Learning Framework

SafetyDGX agent

arXiv:2605.20373v1 Announce Type: cross Abstract: Building humanoid robots capable of generalizable whole-body loco-manipulation in the real world remains a fundamental challenge. Existing methods eit

Supervised Latent Restructuring for Small-Data Quantum Learning in Plant Phenomics

SafetyDGX agent

arXiv:2605.20413v1 Announce Type: new Abstract: High-dimensional biological data often exhibit a severe mismatch between feature dimensionality and sample size, making reliable classification difficul

SURF: Steering the Scalarization Weight to Uniformly Traverse the Pareto Front

SafetyDGX agent

arXiv:2605.20619v1 Announce Type: new Abstract: Scalarization is widely used in multi-objective optimization owing to its simplicity and scalability. In many applications, the goal is to generate solu

SynCB: A Synergy Concept-Based Model with Dynamic Routing Between Concepts and Complementary Neural Branches

SafetyDGX agent

arXiv:2605.20908v1 Announce Type: new Abstract: Concept-based (CB) models provide interpretability and support test-time human intervention, while standard neural networks (NN) offer strong task perfo

Synchronization and Turn-Taking in Full-Duplex Speech Dialogue Models

SafetyDGX agent

arXiv:2605.20356v1 Announce Type: new Abstract: Full-duplex spoken dialogue models (SDMs) can listen and speak simultaneously, enabling interaction dynamics closer to human conversation than turn-base

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

SafetyDGX agent

arXiv:2506.08277v3 Announce Type: replace-cross Abstract: Recent voxel-wise multimodal brain encoding studies have shown that multimodal large language models (MLLMs) exhibit a higher degree of brain

TelePhysics: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction

SafetyDGX agent

arXiv:2605.20290v1 Announce Type: cross Abstract: Recent generative video models achieve impressive visual quality but remain constrained by limited physical consistency and controllability. Existing

The Economics of AI Inference: Inflation Dynamics, Welfare Costs, and Optimal Monetary Policy under the Inference-Cost Phillips Curve

SafetyDGX agent

arXiv:2605.20281v1 Announce Type: cross Abstract: We develop a unified microeconomic and monetary theory of artificial intelligence inference costs and their pass-through to inflation, welfare, and op

The Illusion of Intervention: Your LLM-Simulated Experiment is an Observational Study

SafetyDGX agent

arXiv:2605.20767v1 Announce Type: new Abstract: Large language models (LLMs) show potential as simulators of human behavior, offering a scalable way to study responses to interventions. However, becau

Time-Prompt: Integrated Heterogeneous Prompts for Unlocking LLMs in Time Series Forecasting

SafetyDGX agent

arXiv:2506.17631v4 Announce Type: replace Abstract: Time series forecasting aims to model temporal dependencies among variables for future state inference, holding significant importance and widesprea

TRAM: Test-Time Risk Adaptation with Mixture of Agents

Model ReleasesDGX agent

arXiv:2408.08812v2 Announce Type: replace Abstract: Deployed reinforcement learning agents often face safety requirements that are specified only after training, such as new hazard maps, revised risk

Tutor-Student Reinforcement Learning: A Dynamic Curriculum for Robust Deepfake Detection

SafetyDGX agent

arXiv:2603.24139v2 Announce Type: replace Abstract: Standard supervised training for deepfake detection treats all samples with uniform importance, which can be suboptimal for learning robust and gene

Uh-oh, the International Space Station is leaking again

SafetyDGX agent

The International Space Station has an ongoing air leak in its Russian PrK module that has been known since 2019 , and this leak has delayed the launch of Axiom Space's private astronaut mission . NAS

Uncertainty-Calibrated Explainable Artificial Intelligence for Fetal Ultrasound Plane Classification: A Systematic Review

SafetyDGX agent

arXiv:2601.00990v2 Announce Type: replace-cross Abstract: Fetal ultrasound is the cornerstone of antenatal care, and accurate recognition of a small set of standard anatomical planes underpins biometr

Velocityformer: Broken-Symmetry-Matched Equivariant Graph Transformers for Cosmological Velocity Reconstruction

SafetyDGX agent

arXiv:2605.21483v1 Announce Type: cross Abstract: Precise measurement of the kinematic Sunyaev-Zel'dovich (kSZ) effect - a probe of the large-scale distribution of baryonic matter, a key observable fo

Verifiable Error Bounds for Physics-Informed Neural Network Solutions of Lyapunov and Hamilton-Jacobi-Bellman Equations

SafetyDGX agent

arXiv:2603.19545v2 Announce Type: replace-cross Abstract: Many core problems in nonlinear systems analysis and control can be recast as solving partial differential equations (PDEs) such as Lyapunov a

VLANeXt: Recipes for Building Strong VLA Models

SafetyDGX agent

arXiv:2602.18532v2 Announce Type: replace Abstract: Following the rise of large foundation models, Vision-Language-Action models (VLAs) emerged, leveraging strong visual and language understanding fro

Waymo suspends freeway rides and pauses its Atlanta operations, as it updates software to improve performance around construction zones and flooded roadways (Reuters)

SafetyDGX agent

Reuters: Waymo suspends freeway rides and pauses its Atlanta operations, as it updates software to improve performance around construction zones and flooded roadways — Alphabet's (GOOGL.O) Waymo said

What Semantics Survive the Connector? Diagnosing VLM-to-DiT Alignment in Video Editing

SafetyDGX agent

arXiv:2605.20795v1 Announce Type: new Abstract: Flow matching based video generative models have been increasingly relying on prepended Vision-Language Models (VLMs) to handle complex, instruction-bas

When does circular financing stop? Similar to a Ponzi Scheme, when cash outflows become greater than inflows. In this last quarter, $NVDA ca…

SafetyDGX agent

When does circular financing stop? Similar to a Ponzi Scheme, when cash outflows become greater than inflows. In this last quarter, NVDA cash grew by only ~600, and now ~95% of its operating cash flow

Why Aggregate Accuracy is Inadequate for Evaluating Fairness in Law Enforcement Facial Recognition Systems

SafetyDGX agent

arXiv:2603.28675v2 Announce Type: replace Abstract: Facial recognition systems are increasingly deployed in law enforcement and security contexts, where algorithmic decisions can carry significant soc

Why Ask One When You Can Ask k? Learning-to-Defer to the Top-k Experts

SafetyDGX agent

arXiv:2504.12988v5 Announce Type: replace Abstract: Existing Learning-to-Defer (L2D) frameworks are limited to single-expert deferral, forcing each query to rely on only one expert and preventing the

Wonder what fraction of those liking this realized it was a joke 🤣

SafetyDGX agent

Wonder what fraction of those liking this realized it was a joke 🤣 There's a lot of excitement about OpenAI's resolution of the unit distance conjecture. In this thread I will discuss the striking app

You can be a cheerleader, or you can be a scientist. Can’t be both. cc: @scaling01

SafetyDGX agent

Gary Marcus argues that pursuing cheerleading and science as careers are mutually exclusive paths, suggesting fundamental incompatibility between these pursuits. The post likely critiques the idea of

Yowza! Similar in some ways to the Stanford paper recently on LLMs hallucinating responses to images they never saw.

SafetyDGX agent

Yowza! Similar in some ways to the Stanford paper recently on LLMs hallucinating responses to images they never saw. This was an amazing and incredibly damning experiment using Microsoft Copilot, by @

20 May 2026

A Geometric Analysis of Sign-Magnitude Asymmetry in a ReLU + RMSNorm Block under Ternary Quantization

SafetyDGX agent

arXiv:2605.18933v1 Announce Type: new Abstract: Pre-norm Transformers with RMSNorm tolerate ternary {-1,0,+1} weight quantization with surprisingly small loss (Ma et al., 2024). We give a geometric ex

A Heuristic Approach for Performance Tuning in RL-based Quadrotor Control via Reward Design and Termination Conditions

SafetyDGX agent

arXiv:2605.19166v1 Announce Type: cross Abstract: Reinforcement learning (RL)-based quadrotor control policies have achieved impressive performance in tasks such as fast navigation in cluttered enviro

A Unified Framework for Structure-Aware Clustering and Heterogeneous Causal Graph Learning

SafetyDGX agent

arXiv:2605.19313v1 Announce Type: cross Abstract: In complex multivariate systems, interactions among variables are defined by dependency structures, often encoded as directed acyclic graphs (ext{DAGs

Accurate Evaluation of Quickest Changepoint Detectors via Non-parametric Survival Analysis

SafetyDGX agent

arXiv:2605.18798v1 Announce Type: new Abstract: We propose non-parametric estimators for the average run length (ARL) and average detection delay (ADD) in quickest changepoint detection (QCD) under fi

Aerial Inspection Behaviors via RL-based Quadrotor Control for Under-canopy Forest Environments

SafetyDGX agent

arXiv:2605.19202v1 Announce Type: cross Abstract: This paper addresses the problem of using a deep Reinforcement Learning (RL)-based low-level Quadrotor controller within an autonomous Quadrotor navig

Agent Sandbox on GKE is now available for everyone, and a first look at Agent Substrate

SafetyDGX agent

In just a short time, we’ve seen AI transition from simple chat interfaces to autonomous agents capable of function calling, code execution, and persistent terminal use. But to orchestrate these capab

Atomistic Modeling of Chemical Disorder in Materials: Bridging Classical Methods and AI-Assisted Approaches

SafetyDGX agent

arXiv:2605.19124v1 Announce Type: cross Abstract: Chemical disorder, originating from the mixed occupation of crystallographic sites by multiple elements, is widespread in alloys, ceramics, and compos

Automatically Improving Simulation Physics for Articulated Objects

SafetyDGX agent

arXiv:2605.19136v1 Announce Type: new Abstract: Simulation is a central tool for scalable robot learning, but its effectiveness depends on the quality of object assets. While modern 3D datasets provid

B-cos GNNs: Faithful Explanations through Dynamic Linearity

SafetyDGX agent

arXiv:2605.19778v1 Announce Type: new Abstract: We introduce B-cos GNNs, an inherently explainable class of graph neural networks whose predictions decompose exactly into per-node, per-feature contrib

BERTO: Intent-Driven Network Time Series Forecasting via Natural Language Operator Preferences

SafetyDGX agent

arXiv:2512.05721v2 Announce Type: replace Abstract: Traditional cellular traffic forecasting models are optimized for minimizing symmetric errors, leaving them indifferent to shifting operational prio

Beyond Action Residuals: Real-World Robot Policy Steering via Bottleneck Latent Reinforcement Learning

SafetyDGX agent

arXiv:2605.19919v1 Announce Type: new Abstract: Pretrained imitation policies have become a strong foundation for robot manipulation, but they often require online improvement to overcome execution er

Beyond Extrapolation: Knowledge Utilization Paradigm with Bidirectional Inspiration for Time Series Forecasting

SafetyDGX agent

arXiv:2605.19249v1 Announce Type: new Abstract: Time-series forecasting is critical in various scenarios, such as energy, transportation, and public health. However, most existing forecasters rely pri

Beyond Isotropy in JEPAs: Hamiltonian Geometry and Symplectic Prediction

SafetyDGX agent

arXiv:2605.20107v1 Announce Type: cross Abstract: JEPAs often regularize one-view embeddings toward an isotropic Gaussian, implicitly baking Euclidean symmetry into the representation. We show that th

Beyond Mode Collapse: Distribution Matching for Diverse Reasoning

SafetyDGX agent

arXiv:2605.19461v1 Announce Type: new Abstract: On-policy reinforcement learning methods like GRPO suffer from mode collapse: they exhibit reduced solution diversity, concentrating probability mass on

Boosting Text-to-Image Diffusion Models via Core Token Attention-Based Seed Selection

SafetyDGX agent

arXiv:2605.19532v1 Announce Type: new Abstract: Text-to-image diffusion models can synthesize high-quality images, yet the outcome is notoriously sensitive to the random seed: different initial seeds

Brain alignment of reasoning and action representations from vision-language and action models during naturalistic gameplay

SafetyDGX agent

arXiv:2605.19352v1 Announce Type: cross Abstract: Understanding how humans and artificial intelligence systems predict and plan by interacting with their environment is a fundamental challenge at the

Can Large Language Models Revolutionize Survey Research? Experiments with Disaster Preparedness Responses

SafetyDGX agent

arXiv:2605.19229v1 Announce Type: new Abstract: Survey research faces mounting structural challenges: declining response rates, sample bias, block-wise missingness among at-risk respondents, and AI-as

← Previous
1…155156157158159…242
Next →