AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Don't Retrain, Align: Adapting Autoregressive LMs to Diffusion LMs via Representation Alignment

DGX agent

arXiv:2605.06885v1 Announce Type: cross Abstract: Diffusion language models (DLMs) have recently demonstrated capabilities that complement standard autoregressive (AR) models, particularly in non-sequ

safetyarxiv-cs-ai
11 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Dr-BA: Separable Optimization for Direct Radar Bundle Adjustment & Localization

DGX agent

arXiv:2605.07041v1 Announce Type: cross Abstract: This paper introduces Dr-BA, a first-of-its-kind radar bundle adjustment (BA) framework that operates directly on 2D spinning radar intensity images.

safetyarxiv-cs-cv
11 May 2026
Safety

Dr. Post-Training: A Data Regularization Perspective on LLM Post-Training

DGX agent

arXiv:2605.07063v1 Announce Type: cross Abstract: Data selection methods address a critical challenge in LLM post-training: effectively leveraging scarce, high-fidelity target data alongside abundant

safetyarxiv-cs-ai
11 May 2026
Safety

DReS: Dual Reconstruction Smoothing for Functional Regularization

DGX agent

arXiv:2510.00253v2 Announce Type: replace Abstract: Smoothness is a key inductive bias in machine learning and is closely related to generalization. Existing smoothness-inducing methods typically rely

safetyarxiv-cs-lg
11 May 2026
Safety

Drifting Field Policy: A One-Step Generative Policy via Wasserstein Gradient Flow

DGX agent

arXiv:2605.07727v1 Announce Type: cross Abstract: We propose Drifting Field Policy (DFP), a non-ODE one-step generative policy built on the drifting model paradigm. We frame the policy update as a rev

safetyarxiv-cs-ai
11 May 2026
Safety

Dynamic one-time delivery of critical data by small and sparse UAV swarms: a model problem for MARL scaling studies

DGX agent

arXiv:2512.09682v2 Announce Type: replace-cross Abstract: This work studies the application of Multi-Agent Reinforcement Learning (MARL) to decentralized control of unmanned aerial vehicles to relay a

safetyarxiv-cs-ai
11 May 2026
Safety

Econometric vs. Causal Structure-Learning for Time-Series Policy Decisions: Evidence from the UK COVID-19 Policies

DGX agent

arXiv:2603.00041v2 Announce Type: replace-cross Abstract: Causal machine learning (ML) recovers graphical structures that inform us about potential cause-and-effect relationships. Most progress has fo

safetyarxiv-cs-ai
11 May 2026
Safety

EditTransfer++: Toward Faithful and Efficient Visual-Prompt-Guided Image Editing

DGX agent

arXiv:2605.07455v1 Announce Type: new Abstract: Visual-prompt-guided edit transfer aims to learn image transformations directly from example pairs, offering more precise and controllable editing than

safetyarxiv-cs-cv
11 May 2026
Safety

EGA: Adapting Frozen Encoders for Vector Search with Bounded Out-of-Distribution Degradation

DGX agent

arXiv:2605.05674v2 Announce Type: replace-cross Abstract: Vector search systems built on frozen vision encoders face queries from unseen classes at deployment, yet existing adapter training collapses

safetyarxiv-cs-ai
11 May 2026
Safety

Emergent social transmission of model-based representations without inference

DGX agent

arXiv:2604.05777v2 Announce Type: replace Abstract: How do people acquire rich, flexible knowledge about their environment from others despite limited cognitive capacity? Humans are often thought to r

safetyarxiv-cs-ai
11 May 2026
Safety

Emergent Symbolic Structure in Health Foundation Models: Extraction, Alignment, and Cross-Modal Transfer

DGX agent

arXiv:2605.07407v1 Announce Type: new Abstract: Health foundation models (FMs) learn useful representations from wearable sensors, but interpreting what they encode and transferring that knowledge acr

safetyarxiv-cs-lg
11 May 2026
Safety

Entropy-Regularized Adjoint Matching for Offline Reinforcement Learning

DGX agent

arXiv:2605.06156v2 Announce Type: replace-cross Abstract: Integrating expressive generative policies, such as flow-matching models, into offline reinforcement learning (RL) allows agents to capture co

safetyarxiv-cs-ai
11 May 2026
Safety

Escaping the Diversity Trap in Robotic Manipulation via Anchor-Centric Adaptation

DGX agent

arXiv:2605.07381v1 Announce Type: cross Abstract: While Vision-Language-Action (VLA) models offer broad general capabilities, deploying them on specific hardware requires real-world adaptation to brid

safetyarxiv-cs-ai
11 May 2026
Safety

EvolveR: Self-Evolving LLM Agents through an Experience-Driven Lifecycle

DGX agent

arXiv:2510.16079v2 Announce Type: replace-cross Abstract: Current Large Language Model (LLM) agents show strong performance in tool use, but lack the crucial capability to systematically learn from th

safetyarxiv-cs-ai
11 May 2026
Safety

Experience Sharing in Mutual Reinforcement Learning for Heterogeneous Language Models

DGX agent

arXiv:2605.07244v1 Announce Type: cross Abstract: We introduce Mutual Reinforcement Learning, a framework for concurrent RL post-training in which heterogeneous LLM policies exchange typed experience

safetyarxiv-cs-ai
11 May 2026
Safety

f-Divergence Regularized RLHF: Two Tales of Sampling and Unified Analyses

DGX agent

arXiv:2605.06977v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) has become a cornerstone technique for post-training large language models. While most existing appr

safetyarxiv-cs-ai
11 May 2026
Safety

Flatness and Gradient Alignment Are Both Necessary: Spectral-Aware Gradient-Aligned Exploration for Multi-Distribution Learning

DGX agent

arXiv:2605.07914v1 Announce Type: cross Abstract: Sharpness-aware and gradient-alignment methods have been shown to improve generalization, however each family of methods targets a single geometric pr

safetyarxiv-cs-cv
11 May 2026
Safety

Flexible Entropy Control in RLVR with a Gradient-Preserving Perspective

DGX agent

arXiv:2602.09782v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a critical method for enhancing the reasoning capabilities of Large Langu

safetyarxiv-cs-ai
11 May 2026
Safety

Flow-OPD: On-Policy Distillation for Flow Matching Models

DGX agent

arXiv:2605.08063v1 Announce Type: cross Abstract: Existing Flow Matching (FM) text-to-image models suffer from two critical bottlenecks under multi-task alignment: the reward sparsity induced by scala

safetyarxiv-cs-ai
11 May 2026
Safety

From 0-Order Selection to 2-Order Judgment: Combinatorial Hardening Exposes Compositional Failures in Frontier LLMs

DGX agent

arXiv:2605.07268v1 Announce Type: new Abstract: Multiple-choice reasoning benchmarks face dual challenges: rapid saturation from advancing models and data contamination that undermines static evaluati

safetyarxiv-cs-cl
11 May 2026
Safety

From Model to Data (M2D): Shifting Complexity from GNNs to Graphs for Transparent Graph Learning

DGX agent

arXiv:2605.06814v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) achieve high performance but can be opaque to humans, making it difficult to understand and compare the many proposed archi

safetyarxiv-cs-lg
11 May 2026
Safety

From Specification to Deployment: Empirical Evidence from a W3C VC + DID Trust Infrastructure for Autonomous Agents

DGX agent

arXiv:2605.06738v1 Announce Type: cross Abstract: Autonomous AI agents now transact at production scale -- 69,000 bots executing 165 million transactions across 50 million USDC in cumulative volume on

safetyarxiv-cs-ai
11 May 2026
Safety

From Time Series Analysis to Question Answering: A Survey in the LLM Era

DGX agent

arXiv:2506.11512v2 Announce Type: replace-cross Abstract: Recently, Large Language Models (LLMs) have introduced a novel paradigm in Time Series Analysis (TSA), leveraging strong language capabilities

safetyarxiv-cs-ai
11 May 2026
Safety

GASim: A Graph-Accelerated Hybrid Framework for Social Simulation

DGX agent

arXiv:2605.07692v1 Announce Type: new Abstract: Large-scale social simulators are essential for studying complex social patterns. Prior work explores hybrid methods to scale up simulations, combining

safetyarxiv-cs-ai
11 May 2026
Safety

Generalised Linear Models in Deep Bayesian RL with Learnable Basis Functions

DGX agent

arXiv:2512.20974v3 Announce Type: replace-cross Abstract: Bayesian Reinforcement Learning (BRL), a subclass of Meta-Reinforcement Learning (Meta-RL), provides a principled framework for generalisation

safetyarxiv-cs-ai
11 May 2026
Safety

GRAPE: Let GRPO Supervise Query Rewriting by Ranking for Retrieval

DGX agent

arXiv:2509.23370v2 Announce Type: replace Abstract: The CLIP model has established itself as a cornerstone of large-scale retrieval systems. However, its performance often degrades under distributiona

safetyarxiv-cs-cv
11 May 2026
Safety

Guidance Is Not a Hyperparameter: Learning Dynamic Control in Diffusion Language Models

DGX agent

arXiv:2605.07701v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used mechanism for controlling diffusion-based generative models, yet its guidance scale is typically treated

safetyarxiv-cs-cl
11 May 2026
Safety

GustPilot: A Hierarchical DRL-INDI Framework for Wind-Resilient Quadrotor Navigation

DGX agent

arXiv:2603.19966v2 Announce Type: replace Abstract: Wind disturbances remain a key barrier to reliable autonomous navigation for lightweight quadrotors, where the rapidly varying airflow can destabili

safetyarxiv-cs-ro
11 May 2026
Safety

HAIC: Humanoid Agile Object Interaction Control via Dynamics-Aware World Model

DGX agent

arXiv:2602.11758v2 Announce Type: replace Abstract: Humanoid robots show promise for complex whole-body tasks in unstructured environments. Although Human-Object Interaction (HOI) has advanced, most m

safetyarxiv-cs-ro
11 May 2026
Safety

HEART: Hyperspherical Embedding Alignment via Kent-Representation Traversal in Diffusion Models

DGX agent

arXiv:2605.07973v1 Announce Type: new Abstract: Text-to-image diffusion models can generate visually stunning images, yet, controlling what appears and how it appears, remains surprisingly difficult,

safetyarxiv-cs-cv
11 May 2026
Safety

How Log-Barrier Helps Exploration in Policy Optimization

DGX agent

arXiv:2603.15001v2 Announce Type: replace-cross Abstract: Recently, it has been shown that the Stochastic Gradient Bandit (SGB) algorithm converges to a globally optimal policy with a constant learnin

safetyarxiv-cs-ai
11 May 2026
Safety

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment

DGX agent

arXiv:2605.06850v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a crucial paradigm for unlocking the advanced reasoning capabilities of Large Language Models (LLMs), encom

safetyarxiv-cs-ai
11 May 2026
Safety

How to utilize failure demo data?: Effective data selection for imitation learning using distribution differences in attention mechanism

DGX agent

arXiv:2605.07560v1 Announce Type: new Abstract: Imitation learning for robotic tasks has relied primarily on policies trained only on successful demonstrations, although failures are unavoidable durin

safetyarxiv-cs-ro
11 May 2026
Safety

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training

DGX agent

arXiv:2605.07316v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning

safetyarxiv-cs-ai
11 May 2026
Safety

Inference-Time Attribute Distribution Alignment for Unconditional Diffusion

DGX agent

arXiv:2605.07456v1 Announce Type: new Abstract: Inference-time controllable generation is essential for real-world applications of unconditional diffusion models. However, most existing techniques foc

safetyarxiv-cs-lg
11 May 2026
Safety

InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization

DGX agent

arXiv:2605.07099v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is fundamental for precise localization and navigation in GPS-denied environments, aiming to match ground or UAV imag

safetyarxiv-cs-cv
11 May 2026
Safety

InterCoG: Towards Spatially Precise Image Editing with Interleaved Chain-of-Grounding Reasoning

DGX agent

arXiv:2603.01586v3 Announce Type: replace Abstract: Emerging unified editing models have demonstrated strong capabilities in general object editing tasks. However, it remains a significant challenge t

safetyarxiv-cs-cv
11 May 2026
Safety

Inverse Reinforcement Learning with Just Classification and a Few Regressions

DGX agent

arXiv:2509.21172v2 Announce Type: replace Abstract: Inverse reinforcement learning (IRL) aims to infer rewards from observed behavior, but rewards are not identified from the policy alone: many reward

safetyarxiv-cs-lg
11 May 2026
Safety

Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models

DGX agent

arXiv:2605.07514v1 Announce Type: cross Abstract: World Action Models (WAMs) enable decision-making through imagined rollouts by predicting future observations and actions. However, the reliability of

safetyarxiv-cs-cv
11 May 2026
Safety

KL for a KL: On-Policy Distillation with Control Variate Baseline

DGX agent

arXiv:2605.07865v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) has emerged as a dominant post-training paradigm for large language models, especially for reasoning domains. However, OP

safetyarxiv-cs-ai
11 May 2026
Safety

Learning Cross-Atlas Consistent Brain Disorder Representations via Disentangled Multi-Atlas Functional Connectivity Learning

DGX agent

arXiv:2605.07026v1 Announce Type: cross Abstract: Functional connectivity (FC) derived from resting-state fMRI is widely used to characterize large-scale brain network alterations in neurological and

safetyarxiv-cs-ai
11 May 2026
Safety

Learning to Track Instance from Single Nature Language Description

DGX agent

arXiv:2605.07064v1 Announce Type: new Abstract: How to achieve vision-language (VL) tracking using natural language descriptions from a video sequence extbf{without relying on any bounding-box ground

safetyarxiv-cs-cv
11 May 2026
Safety

Learning Visual Feature-Based World Models via Residual Latent Action

DGX agent

arXiv:2605.07079v1 Announce Type: cross Abstract: World models predict future transitions from observations and actions. Existing works predominantly focus on image generation only. Visual feature-bas

safetyarxiv-cs-ai
11 May 2026
Safety

Lightweight Unpaired Smartphone ISP Transfer with Semantic Pseudo-Pairing

DGX agent

arXiv:2605.07495v1 Announce Type: new Abstract: Unpaired smartphone ISP is a challenging problem due to the lack of scene and color alignment between RAW and target RGB images. Many existing methods e

safetyarxiv-cs-cv
11 May 2026
Safety

MAGIQ: A Post-Quantum Multi-Agentic AI Governance System with Provable Security

DGX agent

arXiv:2605.06933v1 Announce Type: new Abstract: Our computing ecosystem is being transformed by two emerging paradigms: the increased deployment of agentic AI systems and advancements in quantum compu

safetyarxiv-cs-lg
11 May 2026
Safety

MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge

DGX agent

arXiv:2507.21183v5 Announce Type: replace-cross Abstract: As the era of large language models (LLMs) unfolds, Preference Optimization (PO) methods have become a central approach to aligning LLMs with

safetyarxiv-cs-ai
11 May 2026
Safety

Masks Can Talk: Extracting Structured Text Information from Single-Modal Images for Remote Sensing Change Detection

DGX agent

arXiv:2605.07178v1 Announce Type: new Abstract: Remote sensing change detection is pivotal for urban monitoring, disaster assessment, and environmental resource management. Yet, unimodal deep learning

safetyarxiv-cs-cv
11 May 2026
Safety

Mind the Gap: Geometrically Accurate Generative Reconstruction from Disjoint Views

DGX agent

arXiv:2605.07550v1 Announce Type: new Abstract: 3D vision systems are fundamentally constrained by their reliance on visual overlap: reconstruction methods require it for geometric alignment, while ge

safetyarxiv-cs-cv
11 May 2026
← Previous
1…193194195196197…257
Next →