AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Cross-lingual robustness of LLM-brain alignment and its computational roots

DGX agent

arXiv:2605.21049v1 Announce Type: new Abstract: Large language models (LLMs) reliably predict neural activity during language comprehension and transformer depth has been interpreted as mirroring hier

safetyarxiv-cs-cl
21 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Cumulative Meta-Learning from Active Learning Queries for Robustness to Spurious Correlations

DGX agent

arXiv:2605.20771v1 Announce Type: new Abstract: Spurious correlations in real-world datasets cause machine learning models to rely on irrelevant patterns, undermining reliability, generalization, and

safetyarxiv-cs-lg
21 May 2026
Safety

Data-Efficient Neural Operator Training via Physics-Based Active Learning

DGX agent

arXiv:2605.21348v1 Announce Type: new Abstract: Solving partial differential equations with neural operators significantly reduces computational costs but remains bottlenecked by high training data re

safetyarxiv-cs-lg
21 May 2026
Safety

Decision-Path Patterns as Tree Reliability Signals: Path-based Adaptive Weighting for Random Forest Classification

DGX agent

arXiv:2605.20716v1 Announce Type: new Abstract: Random forests aggregate tree votes by simple majority, treating all trees as equally informative. We observe that the topological pattern along each tr

safetyarxiv-cs-lg
21 May 2026
Safety

Decomposing MXFP4 quantization error for LLM reinforcement learning: reducible bias, recoverable deadzone, and an irreducible floor

DGX agent

arXiv:2605.20402v1 Announce Type: new Abstract: MXFP4 arithmetic can dramatically accelerate reinforcement learning (RL) post-training of large language models (LLMs), yet the quantization error intro

safetyarxiv-cs-lg
21 May 2026
Safety

DeCoR: Design and Control Co-Optimization for Urban Streets Using Reinforcement Learning

DGX agent

arXiv:2605.21311v1 Announce Type: new Abstract: Modern vision systems can detect, track, and forecast urban actors at scale, yet translating perception outputs to urban design remains limited. We intr

safetyarxiv-cs-lg
21 May 2026
Safety

Decoupling Communication from Policy: Robust MARL under Bandwidth Constraints

DGX agent

arXiv:2605.21085v1 Announce Type: cross Abstract: Communication enables coordination in multi-agent reinforcement learning (MARL), but many real-world applications, e.g., search-and-rescue with drone

safetyarxiv-cs-lg
21 May 2026
Safety

Deep Attention Reweighting: Post-Hoc Attention-Based Feature Aggregation in CNNs for Disentangling Core and Spurious Features under Spurious Correlations

DGX agent

arXiv:2605.20732v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) often exploit spurious correlations in datasets, learning superficially predictive yet causally irrelevant features

safetyarxiv-cs-cv
21 May 2026
Safety

DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards

DGX agent

arXiv:2605.21467v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards (RLVR) has emerged as a central technique for improving the reasoning capabilities of large language mo

safetyarxiv-cs-cl
21 May 2026
Safety

Design for Manufacturing: A Manufacturability Knowledge-Integrated Reinforcement Learning Framework for Free-Form Pipe Routing in Aeroengines

DGX agent

arXiv:2605.20644v1 Announce Type: new Abstract: Design for manufacturing plays a critical role in advanced aeroengine development, where complex components necessitate careful consideration of manufac

safetyarxiv-cs-lg
21 May 2026
Safety

Disentangling Bias by Modeling Intra- and Inter-modal Causal Attention for Multimodal Sentiment Analysis

DGX agent

arXiv:2508.04999v2 Announce Type: replace Abstract: Multimodal sentiment analysis (MSA) aims to understand human emotions by integrating information from multiple modalities, such as text, audio, and

safetyarxiv-cs-lg
21 May 2026
Safety

Distributed Direct Preference Optimization

DGX agent

arXiv:2605.20696v1 Announce Type: new Abstract: Preference-based reinforcement learning (RL) is a key paradigm for aligning policies with human judgments, yet its theoretical behavior in distributed s

safetyarxiv-cs-lg
21 May 2026
Safety

Distribution-Aware Reward: Reinforcement Learning over Predictive Distributions for LLM Regression

DGX agent

arXiv:2605.20740v1 Announce Type: cross Abstract: Large language models can predict real-valued quantities from heterogeneous inputs such as text, code, and molecular strings, but most training object

safetyarxiv-cs-cl
21 May 2026
Safety

Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning

DGX agent

arXiv:2605.20730v1 Announce Type: new Abstract: In-context learning (ICL) allows large language models (LLMs) to adapt to new tasks through demonstrations, yet it suffers from escalating inference cos

safetyarxiv-cs-cl
21 May 2026
Safety

Divide-Prompt-Refine: a Training-Free, Structure-Aware Framework for Biomedical Abstract Generation

DGX agent

arXiv:2605.20628v1 Announce Type: new Abstract: Biomedical abstracts play a critical role in downstream NLP applications, such as information retrieval, biocuration, and biomedical knowledge discovery

safetyarxiv-cs-cl
21 May 2026
Safety

Enhancing Speech Large Language Models through Reinforced Behavior Alignment

DGX agent

arXiv:2509.03526v2 Announce Type: replace Abstract: The recent advancements of Large Language Models (LLMs) have spurred considerable research interest in extending their linguistic capabilities beyon

safetyarxiv-cs-cl
21 May 2026
Safety

EvalMORAAL: Interpretable Chain-of-Thought and LLM-as-Judge Evaluation for Moral Alignment in Large Language Models

DGX agent

arXiv:2510.05942v3 Announce Type: replace Abstract: We present EvalMORAAL, a transparent chain-of-thought (CoT) framework that uses two scoring methods (log-probabilities and direct ratings) plus a mo

safetyarxiv-cs-cl
21 May 2026
Safety

extit{Stochastic} MeanFlow Policies: One-Step Generative Control with Entropic Mirror Descent

DGX agent

arXiv:2605.21282v1 Announce Type: new Abstract: Online off-policy reinforcement learning (RL) is shaped by two coupled choices: the policy class and the update rule. Gaussian policies are fast and hav

safetyarxiv-cs-lg
21 May 2026
Safety

FEAT: A Linear-Complexity Foundation Model for Extremely Large Structured Data

DGX agent

arXiv:2603.16513v3 Announce Type: replace Abstract: Structured data is widely used in domains such as healthcare, finance, and scientific data management. Recent studies on structured data foundation

safetyarxiv-cs-lg
21 May 2026
Safety

FlowLong: Inference-time Long Video Generation via Manifold-constrained Tweedie Matching

DGX agent

arXiv:2605.20910v1 Announce Type: new Abstract: Extending the generation horizon of video diffusion models to long sequences remains a long-standing and important challenge. Existing training-free app

safetyarxiv-cs-cv
21 May 2026
Safety

Graph Transductive Sharpening: Leveraging Unlabeled Predictions in Node Classification

DGX agent

arXiv:2605.20248v1 Announce Type: new Abstract: In the transductive setting, where the full graph is observed but node labels are only partially available, progress in semi-supervised node classificat

safetyarxiv-cs-lg
21 May 2026
Safety

GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents

DGX agent

arXiv:2605.20246v1 Announce Type: new Abstract: Recently, vision-language model (VLM) agents have shown promising progress in open-world tasks, where successful task completion often requires multiple

safetyarxiv-cs-lg
21 May 2026
Safety

HORST: Composing Optimizer Geometries for Sparse Transformer Training

DGX agent

arXiv:2605.21104v1 Announce Type: new Abstract: Sparsifying transformers remains a fundamental challenge, as standard optimizers fail to simultaneously encourage sparsity and maintain training stabili

safetyarxiv-cs-lg
21 May 2026
Safety

Improved convergence rate of kNN graph Laplacians: differentiable self-tuned affinity

DGX agent

arXiv:2410.23212v2 Announce Type: replace-cross Abstract: In graph-based data analysis, k-nearest neighbor (kNN) graphs are widely used due to their adaptivity to local data densities. Allowing weight

safetyarxiv-cs-lg
21 May 2026
Safety

Inference Time Policy Optimization for Offline RL with Differentiable World Models

DGX agent

arXiv:2603.22430v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) learns optimal policies from fixed datasets, training a policy once and deploying it at inference time without f

safetyarxiv-cs-lg
21 May 2026
Safety

It Takes Two: Complementary Self-Distillation for Contextual Integrity in LLMs

DGX agent

arXiv:2605.20258v1 Announce Type: new Abstract: Contextual Integrity (CI) defines privacy not merely as keeping information hidden, but as governing information flows according to the norms of a given

safetyarxiv-cs-lg
21 May 2026
Safety

Latent Geometry as a Structural Monitor: Eigenspace Alignment for Anomaly Detection in Anonymity Networks

DGX agent

arXiv:2605.20391v1 Announce Type: cross Abstract: Traditional anomaly detection marks events when measured signals cross predefined thresholds. This captures the moment of transition but not the struc

safetyarxiv-cs-lg
21 May 2026
Safety

Learning Robust Dexterous In-Hand Manipulation from Joint Sensors with Proprioceptive Transformer

DGX agent

arXiv:2605.21330v1 Announce Type: new Abstract: In-hand object manipulation is a fundamental yet challenging capability for dexterous robots. Despite significant progress in dexterous manipulation, ex

safetyarxiv-cs-ro
21 May 2026
Safety

Learning to Think in Physics: Breaking Shortcut Learning in Scientific Diffusion via Representation Alignment

DGX agent

arXiv:2605.20780v1 Announce Type: cross Abstract: Physics-informed diffusion models typically enforce PDE constraints only on final outputs, leaving intermediate representations unconstrained and pron

safetyarxiv-cs-cv
21 May 2026
Safety

Letting Trajectories Spread: Quality-Preserving Control for Diverse Flow Matching

DGX agent

arXiv:2510.09060v2 Announce Type: replace-cross Abstract: Flow-based text-to-image models follow deterministic trajectories, making it costly to explore diverse modes under limited sampling budgets. E

safetyarxiv-cs-cv
21 May 2026
Safety

Linear-DPO: Linear Direct Preference Optimization for Diffusion and Flow-Matching Generative Models

DGX agent

arXiv:2605.21123v1 Announce Type: new Abstract: Direct Preference Optimization (DPO) is successful for alignment in LLMs but still faces challenges in text-to-image generation. Existing studies are co

safetyarxiv-cs-cv
21 May 2026
Safety

Listwise Policy Optimization: Group-based RLVR as Target-Projection on the LLM Response Simplex

DGX agent

arXiv:2605.06139v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) has become a standard approach for large language models (LLMs) post-training to incentivize r

safetyarxiv-cs-lg
21 May 2026
Safety

LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series

DGX agent

arXiv:2605.20449v1 Announce Type: new Abstract: Can language-pretrained transformers become effective time-series forecasters, and why? In this paper, we show that cross-modal transfer arises because

safetyarxiv-cs-lg
21 May 2026
Safety

Mahjax: A GPU-Accelerated Mahjong Simulator for Reinforcement Learning in JAX

DGX agent

arXiv:2605.20577v1 Announce Type: cross Abstract: Riichi Mahjong is a multi-player, imperfect-information game characterized by stochasticity and high-dimensional state spaces. These attributes presen

safetyarxiv-cs-lg
21 May 2026
Safety

Mind the Sim-to-Real Gap & Think Like a Scientist

DGX agent

arXiv:2605.21458v1 Announce Type: cross Abstract: Suppose a planner has a pre-trained simulator of a sequential decision problem and the option to run real experiments in the field. The simulator is c

safetyarxiv-cs-lg
21 May 2026
Safety

Mitigating Label Bias with Interpretable Rubric Embeddings

DGX agent

arXiv:2605.21455v1 Announce Type: new Abstract: Statistical decision algorithms are increasingly deployed in domains where ground-truth labels are hard to obtain, such as hiring, university admissions

safetyarxiv-cs-lg
21 May 2026
Safety

Mobile UMI: Cross-View Diffusion Policy with Decoupled Kinematics for Mobile Manipulation

DGX agent

arXiv:2605.20894v1 Announce Type: new Abstract: Mobile imitation learning on portable demonstration interfaces faces two coupled bottlenecks: locomotion-contaminated action labels and inference-induce

safetyarxiv-cs-ro
21 May 2026
Safety

Multi-Head Attention as Ensemble Nadaraya-Watson Estimation: Variance Reduction, Decorrelation, and Optimal Head Diversity

DGX agent

arXiv:2605.20271v1 Announce Type: cross Abstract: We develop a rigorous statistical theory of multi-head attention (MHA) as an ensemble of Nadaraya-Watson (NW) kernel regression estimators. Building o

safetyarxiv-cs-lg
21 May 2026
Safety

Multi-Step Likelihood-Ratio Correction for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2605.20865v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) plays a pivotal role in improving the reasoning ability of large language models. However, widely

safetyarxiv-cs-lg
21 May 2026
Safety

Multimodal LLMs under Pairwise Modalities

DGX agent

arXiv:2605.21059v1 Announce Type: new Abstract: Despite the impressive results achieved by multimodal large language models (MLLMs), their training typically relies on jointly curated multimodal data,

safetyarxiv-cs-cv
21 May 2026
Safety

NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control

DGX agent

arXiv:2605.20209v1 Announce Type: cross Abstract: Achieving precise, versatile whole-body character control in physics-based animation remains challenging. Recent diffusion-based policies generate ric

safetyarxiv-cs-lg
21 May 2026
Safety

Neural Collapse by Design: Learning Class Prototypes on the Hypersphere

DGX agent

arXiv:2605.20302v1 Announce Type: cross Abstract: Supervised classification has a theoretical optimum, Neural Collapse (NC), yet neither of its two dominant paradigms reaches it in practice. Cross ent

safetyarxiv-cs-cv
21 May 2026
Safety

OcclusionFormer: Arranging Z-Order for Layout-Grounded Image Generation

DGX agent

arXiv:2605.21343v1 Announce Type: new Abstract: Recent layout-to-image models have achieved remarkable progress in spatial controllability. However, they still struggle with inter-object occlusion. Wh

safetyarxiv-cs-cv
21 May 2026
Safety

Parallel LLM Reasoning for Bias-Resilient, Robust Conceptual Abstraction

DGX agent

arXiv:2605.20194v1 Announce Type: new Abstract: Large language models (LLMs) have been increasingly used to analyze text. However, they are often plagued with contextual reasoning limitations when ana

safetyarxiv-cs-cl
21 May 2026
Safety

Pareto-Enhanced Portrait Generation: Vision-Aligned Text Supervision for Alignment, Realism, and Aesthetics

DGX agent

arXiv:2605.20640v1 Announce Type: new Abstract: Text-to-image diffusion models often face a severe trilemma in human portrait generation: text-image alignment, photorealism, and human-perceived aesthe

safetyarxiv-cs-cv
21 May 2026
Safety

PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction

DGX agent

arXiv:2605.21414v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have shown strong potential for general-purpose robotic manipulation by leveraging large pretrained vision-languag

safetyarxiv-cs-cv
21 May 2026
Safety

Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization

DGX agent

arXiv:2510.21583v2 Announce Type: replace Abstract: Recent Progress in post-training flow matching for text-to-image (T2I) generation with Group Relative Policy Optimization (GRPO) has demonstrated st

safetyarxiv-cs-cv
21 May 2026
Safety

Q-SpiRL: Quantum Spiking Reinforcement Learning for Adaptive Robot Navigation

DGX agent

arXiv:2605.20801v1 Announce Type: new Abstract: Adaptive robot navigation in dynamic environments requires policies that can reach the target reliably while producing efficient and stable trajectories

safetyarxiv-cs-ro
21 May 2026
← Previous
1…167168169170171…260
Next →