AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
15 May 2026

V2M-Zero: Zero-Pair Time-Aligned Video-to-Music Generation

SafetyDGX agent

arXiv:2603.11042v2 Announce Type: replace-cross Abstract: Generating music that temporally aligns with video events is challenging for existing text-to-music models, which lack fine-grained temporal c

Very thoughtful post on the role and success of open source, including potential applications for AVs and AI. Long, but worth the read. -- '…

SafetyDGX agent

Very thoughtful post on the role and success of open source, including potential applications for AVs and AI. Long, but worth the read. -- 'Open source is no longer just how good software gets built.

VGGT-Omega

SafetyDGX agent

arXiv:2605.15195v1 Announce Type: new Abstract: Recent feed-forward reconstruction models, such as VGGT, have proven competitive with traditional optimization-based reconstructors while also providing

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Video-OPD: Efficient Post-Training of Multimodal Large Language Models for Temporal Video Grounding via On-Policy Distillation

SafetyDGX agent

arXiv:2602.02994v2 Announce Type: replace Abstract: Reinforcement learning has emerged as a principled post-training paradigm for Temporal Video Grounding (TVG) due to its on-policy optimization, yet

Vision-Core Guided Contrastive Learning for Balanced Multi-modal Prognosis Prediction of Stroke

SafetyDGX agent

arXiv:2605.14710v1 Announce Type: cross Abstract: Deep learning and multi-modal fusion have demonstrated transformative potential in medical diagnosis by integrating diverse data sources. However, acc

Vision-LLMs for Spatiotemporal Traffic Forecasting

SafetyDGX agent

arXiv:2510.11282v2 Announce Type: replace Abstract: Accurate spatiotemporal traffic forecasting is a critical prerequisite for proactive resource management in dense urban mobile networks. While large

Warp-as-History: Generalizable Camera-Controlled Video Generation from One Training Video

SafetyDGX agent

arXiv:2605.15182v1 Announce Type: new Abstract: Camera-controlled video generation has made substantial progress, enabling generated videos to follow prescribed viewpoint trajectories. However, existi

“we are below rock bottom now”. Trump admin continues to piss away America’s scientific lead for no good reason. @elonmusk stands by, does n…

SafetyDGX agent

“we are below rock bottom now”. Trump admin continues to piss away America’s scientific lead for no good reason. @elonmusk stands by, does nothing. NEW: A serious staffing shortage — of the Trump admi

XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations

SafetyDGX agent

arXiv:2511.02776v2 Announce Type: replace Abstract: Recent progress in large-scale robotic datasets and vision-language models (VLMs) has advanced research on vision-language-action (VLA) models. Howe

14 May 2026

140,000 fake citations in 2025 alone😠

SafetyDGX agent

Gary Marcus highlighted a concerning problem in 2025 where approximately 140,000 fabricated citations were identified, likely referring to false or hallucinated references generated by AI systems in a

A_3B_2: Adaptive Asymmetric Adapter for Alleviating Branch Bias in Vision-Language Image Classification with Few-Shot Learning

SafetyDGX agent

arXiv:2605.13161v1 Announce Type: new Abstract: Efficient transfer learning methods for large-scale vision-language models (e.g., CLIP) enable strong few-shot transfer, yet existing adaptation methods

Achieving epsilon^{-2} Sample Complexity for Single-Loop Actor-Critic under Minimal Assumptions

SafetyDGX agent

arXiv:2605.13639v1 Announce Type: new Abstract: In this paper, we establish last-iterate convergence rates for off-policy actor--critic methods in reinforcement learning. In particular, under a single

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

SafetyDGX agent

arXiv:2605.12569v1 Announce Type: cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environ

Adam-SHANG: A Convergent Adam-Type Method for Stochastic Smooth Convex Optimization

SafetyDGX agent

arXiv:2605.12878v1 Announce Type: cross Abstract: We propose Adam-SHANG, a Lyapunov-guided Adam-type method that couples momentum, adaptive preconditioning, and a curvature-aware correction through a

Adaptive mine planning under geological uncertainty: A POMDP framework for sequential decision-making

SafetyDGX agent

arXiv:2605.13702v1 Announce Type: new Abstract: Strategic mine production scheduling under geological uncertainty is conventionally formulated as a stochastic optimization problem in which a fixed ext

Adaptive Smooth Tchebycheff Attention for Multi-Objective Policy Optimization

SafetyDGX agent

arXiv:2605.12771v1 Announce Type: cross Abstract: Multi-objective reinforcement learning in robotic domains requires balancing complex, non-convex trade-offs between conflicting objectives. While line

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

SafetyDGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

Addressing Finite-Horizon MDPs via Low-Rank Tensor Value Approximation

SafetyDGX agent

arXiv:2501.10598v3 Announce Type: replace Abstract: We study the problem of learning optimal policies in finite-horizon Markov Decision Processes (MDPs) using low-rank reinforcement learning (RL) meth

Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding

SafetyDGX agent

arXiv:2602.02977v2 Announce Type: replace-cross Abstract: Vision-language models such as CLIP often struggle to faithfully understand long, detail-rich captions, relying on dominant scene cues while o

Aligning Network Equivariance with Data Symmetry: A Theoretical Framework and Adaptive Approach for Image Restoration

SafetyDGX agent

arXiv:2605.13744v1 Announce Type: new Abstract: Image restoration is an inherently ill posed inverse problem. Equivariant networks that embed geometric symmetry priors can mitigate this ill posedness

AnyFlow: Any-Step Video Diffusion Model with On-Policy Flow Map Distillation

SafetyDGX agent

arXiv:2605.13724v1 Announce Type: cross Abstract: Few-step video generation has been significantly advanced by consistency distillation. However, the performance of consistency-distilled models often

Auditing Sybil: Explaining Deep Lung Cancer Risk Prediction Through Generative Interventional Attributions

SafetyDGX agent

arXiv:2602.02560v2 Announce Type: replace-cross Abstract: Lung cancer remains the leading cause of cancer mortality, driving the development of automated screening tools to alleviate radiologist workl

Beyond Anthropomorphism: Exploring the Roles of Perceived Non-humanity and Structural Similarity in Deep Self-Disclosure Toward Generative AI

SafetyDGX agent

arXiv:2605.13574v1 Announce Type: cross Abstract: This study investigates deep self-disclosure toward generative AI by examining perceived non-humanity and structural similarity as psychological facto

BlitzGS: City-Scale Gaussian Splatting at Lightning Speed

SafetyDGX agent

arXiv:2605.13794v1 Announce Type: cross Abstract: We present BlitzGS, a distributed 3DGS framework that reduces active Gaussian workload for fast city-scale reconstruction. BlitzGS manages this worklo

Block-wise Adaptive Caching for Accelerating Diffusion Policy

SafetyDGX agent

arXiv:2506.13456v2 Announce Type: replace Abstract: Diffusion Policy has demonstrated strong visuomotor modeling capabilities, but its high computational cost renders it impractical for real-time robo

Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.13054v1 Announce Type: cross Abstract: Cross-domain offline reinforcement learning aims to adapt a policy from a source domain to a target domain using only pre-collected datasets, where en

CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding

SafetyDGX agent

arXiv:2605.13544v1 Announce Type: new Abstract: Fine-grained Vision-Language Pre-training (FVLP) demonstrates significant potential in 3D medical image understanding by aligning anatomy-level visual r

Causal Fine-Tuning under Latent Confounded Shift

SafetyDGX agent

arXiv:2410.14375v3 Announce Type: replace Abstract: Adapting to latent confounded shift remains a core challenge in modern AI. This setting is driven by hidden variables that induce spurious correlati

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

SafetyDGX agent

arXiv:2605.13646v1 Announce Type: cross Abstract: End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recentl

Centralized Adaptive Sampling for Reliable Co-Training of Independent Multi-Agent Policies

SafetyDGX agent

arXiv:2508.01049v2 Announce Type: replace Abstract: Independent on-policy policy gradient algorithms are widely used for multi-agent reinforcement learning (MARL) in cooperative and no-conflict games,

Characterizing Universal Object Representations Across Vision Models

SafetyDGX agent

arXiv:2605.13675v1 Announce Type: new Abstract: Deep neural networks trained with different architectures, objectives, and datasets have been reported to converge on similar visual representations. Ho

ChatSR: Multimodal Large Language Models for Scientific Formula Discovery

SafetyDGX agent

arXiv:2406.05410v3 Announce Type: replace Abstract: Current multimodal large language models (MLLMs) are mainly focused on the understanding and processing of perceptual modalities such as images and

CO-MAP: A Reinforcement Learning Approach to the Qubit Allocation Problem

SafetyDGX agent

arXiv:2605.13638v1 Announce Type: cross Abstract: A quantum compiler is a critical piece in the quantum computing pipeline since it allows an abstract quantum circuit to be run on a physical quantum c

Constraints-of-Thought: A Framework for Constrained Reasoning in Language-Model-Guided Search

SafetyDGX agent

arXiv:2510.08992v3 Announce Type: replace Abstract: While researchers have made significant progress in enabling large language models (LLMs) to perform multi-step planning, LLMs struggle to ensure th

Context Matters: Auditing Gender Bias in T2I Generation through Risk-Tiered Use-Case Profiles

SafetyDGX agent

arXiv:2605.13113v1 Announce Type: cross Abstract: Text-to-image (T2I) generative models are increasingly used to produce content for education, media, and public-facing communication, and are starting

Control where your AI agents can browse with Chrome enterprise policies on Amazon Bedrock AgentCore

SafetyDGX agent

In this post, you will configure Chrome enterprise policies to restrict a browser agent to a specific website, observe the policy enforcement through session recording, and demonstrate custom root CA

COSMIC: Concurrent Optimization of Structure, Material, and Integrated Control for robotic systems

SafetyDGX agent

arXiv:2605.12654v1 Announce Type: new Abstract: Replicating and surpassing the autonomy of natural organisms remains a long-standing goal in robotics. Yet most robotic systems have their structure, ma

Counterfactual Reasoning for Causal Responsibility Attribution in Probabilistic Multi-Agent Systems

SafetyDGX agent

arXiv:2605.13077v1 Announce Type: cross Abstract: Responsibility allocation -- determining the extent to which agents are accountable for outcomes -- is a fundamental challenge in the design and analy

CRAFT: Clinical Reward-Aligned Finetuning for Medical Image Synthesis

SafetyDGX agent

arXiv:2605.12650v1 Announce Type: new Abstract: Foundation diffusion models can generate photorealistic natural images, but adapting them to medical imaging remains challenging. In medical adaptation,

CROP: Expert-Aligned Image Cropping via Compositional Reasoning and Optimizing Preference

SafetyDGX agent

arXiv:2605.12545v1 Announce Type: cross Abstract: Aesthetic image cropping aims to enhance the aesthetic quality of an image by improving its composition through spatial cropping. Previous methods oft

Data Agent: Learning to Select Data via End-to-End Dynamic Optimization

SafetyDGX agent

arXiv:2603.07433v2 Announce Type: replace-cross Abstract: Dynamic Data selection aims to accelerate training by prioritizing informative samples during online training. However, existing methods typic

Decision Support for Marketplace Policies under Incomplete Evidence: From Replay to Launch Readiness

SafetyDGX agent

arXiv:2605.12840v1 Announce Type: cross Abstract: Marketplace platforms routinely evaluate pricing and allocation policies using logged observational data, yet strong offline performance does not impl

Decoupling Exploration and Policy Optimization: Uncertainty Guided Tree Search for Hard Exploration

SafetyDGX agent

arXiv:2603.22273v4 Announce Type: replace Abstract: The process of discovery requires active exploration -- the act of collecting new and informative data. However, efficient autonomous exploration re

Delightful Distributed Policy Gradient

SafetyDGX agent

arXiv:2603.20521v2 Announce Type: replace-cross Abstract: Distributed reinforcement learning trains on data from stale, buggy, or mismatched actors, producing actions with high surprisal (negative log

Differentiable Evolutionary Reinforcement Learning

SafetyDGX agent

arXiv:2512.13399v2 Announce Type: replace Abstract: Crafting effective reward signals remains a central challenge in Reinforcement Learning (RL), especially for complex reasoning tasks. Existing autom

Diffusion Model's Generalization Can Be Characterized by Inductive Biases toward a Data-Dependent Ridge Manifold

SafetyDGX agent

arXiv:2602.06021v2 Announce Type: replace-cross Abstract: We study a data-dependent notion of diffusion-model generalization: when a model does not memorize the training set, where do its generated sa

Do Fair Models Reason Fairly? Counterfactual Explanation Consistency for Procedural Fairness in Credit Decisions

SafetyDGX agent

arXiv:2605.12701v1 Announce Type: cross Abstract: Machine learning algorithms in socially sensitive domains (e.g., credit decisions) often focus on equalizing predictive outcomes. However, satisfying

DP-Muon: Differentially Private Optimization via Matrix-Orthogonalized Momentum

SafetyDGX agent

arXiv:2605.12994v1 Announce Type: new Abstract: We study differentially private (DP) training with Muon, a matrix-valued optimizer that updates hidden-layer weights using momentum followed by Newton--

Driving Intents Amplify Planning-Oriented Reinforcement Learning

SafetyDGX agent

arXiv:2605.12625v1 Announce Type: cross Abstract: Continuous-action policies trained on a single demonstrated trajectory per scene suffer from mode collapse: samples cluster around the demonstrated ma

Emergent and Subliminal Misalignment Through the Lens of Data-Mediated Transfer

SafetyDGX agent

arXiv:2605.12798v1 Announce Type: cross Abstract: Fine-tuning LLMs on narrow harmful datasets can induce Emergent Misalignment (EM), where models exhibit misaligned behavior far beyond the fine-tuning

Ergodic Trajectory Design by Learned Pushforward Maps: Provable Coverage via Conditional Flow Matching

SafetyDGX agent

arXiv:2605.13063v1 Announce Type: new Abstract: Designing continuous trajectories whose time-averaged occupancy provably matches a prescribed spatial density (the ergodic coverage problem) is central

ERPPO: Entropy Regularization-based Proximal Policy Optimization

SafetyDGX agent

arXiv:2605.13131v1 Announce Type: new Abstract: Multi-Agent Proximal Policy Optimization (MAPPO) is a variant of the Proximal Policy Optimization (PPO) algorithm, specifically tailored for multi-agent

F-GRPO: Factorized Group-Relative Policy Optimization for Unified Candidate Generation and Ranking

SafetyDGX agent

arXiv:2605.12995v1 Announce Type: new Abstract: Traditional retrieval pipelines optimize utility through stages of candidate retrieval and reranking, where ranking operates over a predefined candidate

Flow Matching for Offline Reinforcement Learning with Discrete Actions

SafetyDGX agent

arXiv:2602.06138v2 Announce Type: replace Abstract: Generative policies based on diffusion models and flow matching have shown strong promise for offline reinforcement learning (RL), but their applica

Follow-Bench: A Unified Motion Planning Benchmark for Socially-Aware Robot Person Following

Model ReleasesDGX agent

arXiv:2509.10796v4 Announce Type: replace Abstract: Robot person following (RPF) -- mobile robots that follow and assist a specific person -- has emerging applications in personal assistance, security

FrameSkip: Learning from Fewer but More Informative Frames in VLA Training

SafetyDGX agent

arXiv:2605.13757v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies are commonly trained from dense robot demonstration trajectories, often collected through teleoperation, by sampli

Frequency Bias and OOD Generalization in Neural Operators under a Variable-Coefficient Wave Equation

SafetyDGX agent

arXiv:2605.12997v1 Announce Type: new Abstract: Neural operators learn to map initial conditions to the terminal solution of partial differential equations (PDEs), providing a surrogate for the full o

GAGPO: Generalized Advantage Grouped Policy Optimization

SafetyDGX agent

arXiv:2605.13217v1 Announce Type: cross Abstract: Reinforcement learning has become a powerful paradigm for post-training large language model agents, yet credit assignment in multi-turn environments

GRACE: Gradient-aligned Reasoning Data Curation for Efficient Post-training

SafetyDGX agent

arXiv:2605.13130v1 Announce Type: new Abstract: Existing reasoning data curation pipelines score whole samples, treating every intermediate step as equally valuable. In reality, steps within a trace c

Graph-Based Financial Fraud Detection with Calibrated Risk Scoring and Structural Regularization

SafetyDGX agent

arXiv:2605.12782v1 Announce Type: new Abstract: Financial transaction fraud prevention faces challenges such as complex relationship structures, concealed behavioral patterns, and dynamically changing

← Previous
1…167168169170171…242
Next →