AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

When Can Digital Personas Reliably Approximate Human Survey Findings?

DGX agent

arXiv:2605.10659v1 Announce Type: cross Abstract: Digital personas powered by Large Language Models (LLMs) are increasingly proposed as substitutes for human survey respondents, yet it remains unclear

safetyarxiv-cs-ai
12 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

When Language Overwrites Vision: Over-Alignment and Geometric Debiasing in Vision-Language Models

DGX agent

arXiv:2605.08245v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) increasingly power high-stakes applications, from medical imaging to autonomous systems, yet they routinely hallucinate,

safetyarxiv-cs-ai
12 May 2026
Safety

When More Parameters Hurt: Foundation Model Priors Amplify Worst-Client Disparity Under Extreme Federated Heterogeneity

DGX agent

arXiv:2605.08992v1 Announce Type: new Abstract: Federated learning (FL) is increasingly used to fine-tune foundation models (FMs) on distributed private data. The community largely assumes that large-

safetyarxiv-cs-lg
12 May 2026
Safety

Where Do Flow Semantics Reside? A Protocol-Native Tabular Pretraining Paradigm for Encrypted Traffic Classification

DGX agent

arXiv:2603.10051v2 Announce Type: replace-cross Abstract: Self-supervised masked modeling shows promise for encrypted traffic classification by masking and reconstructing raw bytes. Yet recent work re

safetyarxiv-cs-ai
12 May 2026
Safety

Why Adam Works Better with eta_1 = eta_2: The Missing Gradient Scale Invariance Principle

DGX agent

arXiv:2601.21739v2 Announce Type: replace-cross Abstract: Adam has been at the core of large-scale training for almost a decade, yet a simple empirical fact remains unaccounted for: both validation sc

safetyarxiv-cs-ai
12 May 2026
Safety

Why Do DiT Editors Drift? Plug-and-Play Low Frequency Alignment in VAE Latent Space

DGX agent

arXiv:2605.08250v1 Announce Type: cross Abstract: Recent advances in diffusion transformers (DiTs) have enabled promising single-turn image editing capabilities. However, multi-turn editing often lead

safetyarxiv-cs-ai
12 May 2026
Safety

WISTERIA: Learning Clinical Representations from Noisy Supervision via Multi-View Consistency in Electronic Health Records

DGX agent

arXiv:2605.09765v1 Announce Type: cross Abstract: Representation learning in electronic health records (EHR) has largely followed paradigms inherited from natural language processing, relying on seque

safetyarxiv-cs-ai
12 May 2026
Safety

X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning

DGX agent

arXiv:2605.05611v2 Announce Type: replace-cross Abstract: In this paper, we present X-Voice, a 0.4B multilingual zero-shot voice cloning model that clones arbitrary voices and enables everyone to spea

safetyarxiv-cs-ai
12 May 2026
Safety

XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies

DGX agent

arXiv:2605.10734v1 Announce Type: new Abstract: For reinforcement learning in the real world online exploration is expensive A common practice in robotic reinforcement learning is to incorporate addit

safetyarxiv-cs-lg
12 May 2026
Safety

A Finite-Iteration Theory for Asynchronous Categorical Distributional Temporal-Difference Learning

DGX agent

arXiv:2605.06866v1 Announce Type: new Abstract: Recent non-asymptotic analyses have substantially advanced the theory of distributional policy evaluation, but they largely concern synchronous full-sta

safetyarxiv-cs-lg
11 May 2026
Safety

A Generalized Singular Value Theory for Neural Networks

DGX agent

arXiv:2605.06938v1 Announce Type: cross Abstract: Building on the abstract Generalized Singular Value Decomposition (GSVD) theory of Brown et al. [2025], we prove that most modern neural architectures

safetyarxiv-cs-ai
11 May 2026
Safety

A Geometric Taxonomy of Hallucinations in LLMs

DGX agent

arXiv:2602.13224v3 Announce Type: replace Abstract: Hallucinations in deployed language models can have real consequences for downstream decisions in domains such as healthcare, legal, and financial s

safetyarxiv-cs-ai
11 May 2026
Safety

A Large-Scale Dataset for Molecular Structure-Language Description via a Rule-Regularized Method

DGX agent

arXiv:2602.02320v3 Announce Type: replace-cross Abstract: Molecular function is largely determined by structure. Accurately aligning molecular structure with natural language is therefore essential fo

safetyarxiv-cs-ai
11 May 2026
Safety

A Statistical Framework for Algorithmic Collective Action with Multiple Collectives

DGX agent

arXiv:2605.06749v1 Announce Type: cross Abstract: As learning systems increasingly shape everyday decisions, Algorithmic Collective Action (ACA), i.e., users coordinating changes to shared data to ste

safetyarxiv-cs-ai
11 May 2026
Safety

Accurate and Efficient Statistical Testing for Word Semantic Breadth

DGX agent

arXiv:2605.08048v1 Announce Type: new Abstract: Measuring the breadth of a word's meaning, or its spread across contexts, has become feasible with contextualized token embeddings. A word type can be r

safetyarxiv-cs-cl
11 May 2026
Safety

Actor-Critic Algorithm for Dynamic Expectile and CVaR

DGX agent

arXiv:2605.07857v1 Announce Type: new Abstract: Optimizing dynamic risk with stochastic policies is challenging in both policy updates and value learning. The former typically requires transition pert

safetyarxiv-cs-lg
11 May 2026
Safety

Actor-Critic with Active Importance Sampling

DGX agent

arXiv:2605.07094v1 Announce Type: new Abstract: This paper introduces the Active-Importance-Sampling Actor-Critic (AISAC) algorithm, an extension of the Actor-Critic framework for reducing variance in

safetyarxiv-cs-lg
11 May 2026
Safety

Adaptive Subspace Projection for Generative Personalization

DGX agent

arXiv:2605.07257v1 Announce Type: new Abstract: Generative personalization often suffers from the semantic collapsing problem (SCP), where a learned personalized concept overpowers the rest of the tex

safetyarxiv-cs-cv
11 May 2026
Safety

Agentic Coding Needs Proactivity, Not Just Autonomy

DGX agent

arXiv:2605.06717v1 Announce Type: cross Abstract: Coding agents are rapidly changing the landscape of software development, moving from inline completion to autonomous systems that edit repositories,

safetyarxiv-cs-ai
11 May 2026
Safety

Anisotropic Modality Align

DGX agent

arXiv:2605.07825v1 Announce Type: cross Abstract: Training multimodal large language models has long been limited by the scarcity of high-quality paired multimodal data. Recent studies show that the s

safetyarxiv-cs-cv
11 May 2026
Safety

APEX: Assumption-free Projection-based Embedding eXamination Metric for Image Quality Assessment

DGX agent

arXiv:2605.07786v1 Announce Type: cross Abstract: As generative models achieve unprecedented visual quality, the gold standard for image evaluation remains traditional feature-distribution metrics (e.

safetyarxiv-cs-ai
11 May 2026
Safety

ART for Diffusion Sampling: A Reinforcement Learning Approach to Timestep Schedule

DGX agent

arXiv:2601.18681v2 Announce Type: replace-cross Abstract: We consider time discretization for score-based diffusion models to generate samples from a learned reverse-time dynamic on a finite grid. Uni

safetyarxiv-cs-ai
11 May 2026
Safety

ASPECT: Node-Level Adaptive Spectral Fusion for Graph Contrastive Learning

DGX agent

arXiv:2604.01878v2 Announce Type: replace-cross Abstract: Spectral graph contrastive learning often constructs low- and high-frequency views to capture complementary graph signals, but these views are

safetyarxiv-cs-ai
11 May 2026
Safety

Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level

DGX agent

arXiv:2605.06387v2 Announce Type: replace-cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories with token-level teacher feedback and often outperforms off-policy disti

safetyarxiv-cs-ai
11 May 2026
Safety

Bellman Calibration for V-Learning in Offline Reinforcement Learning

DGX agent

arXiv:2512.23694v2 Announce Type: replace-cross Abstract: Reliable long-horizon value prediction is difficult in offline reinforcement learning because fitted value methods combine bootstrapping, func

safetyarxiv-cs-lg
11 May 2026
Safety

Better Protein Function Prediction by Modeling Survivorship Bias

DGX agent

arXiv:2605.06879v1 Announce Type: new Abstract: Protein sequence data from nature exhibits survivorship bias: we only observe data from those organisms that survive and reproduce, while non-functional

safetyarxiv-cs-lg
11 May 2026
Safety

Beyond Pairs: Your Language Model is Secretly Optimizing a Preference Graph

DGX agent

arXiv:2605.08037v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) aligns language models using pairwise preference comparisons, offering a simple and effective alternative to Rein

safetyarxiv-cs-ai
11 May 2026
Safety

Beyond State-Wise Mirror Descent: Offline Policy Optimization with Parametric Policies

DGX agent

arXiv:2602.23811v4 Announce Type: replace-cross Abstract: We investigate the theoretical aspects of offline reinforcement learning (RL) under general function approximation. While prior works (e.g., X

safetyarxiv-cs-ai
11 May 2026
Safety

Bias and Uncertainty in LLM-as-a-Judge Estimation

DGX agent

arXiv:2605.06939v1 Announce Type: new Abstract: LLM-as-a-Judge evaluation has become a standard tool for assessing base model performance. However, characterizing performance via the naive estimator,

safetyarxiv-cs-lg
11 May 2026
Safety

CalexNet: Soft Cascade-Aligned Training and Calibration for Lightweight Early-Exit Branches

DGX agent

arXiv:2509.08318v2 Announce Type: replace Abstract: Early-exit cascades over a frozen convolutional backbone enable adaptive inference but suffer from three sources of train-inference mismatch: branch

safetyarxiv-cs-cv
11 May 2026
Safety

Can David Beat Goliath? On Multi-Hop Reasoning with Resource-Constrained Agents

DGX agent

arXiv:2601.21699v2 Announce Type: replace Abstract: Multi-turn reasoning agents solve complex questions by decomposing them into intermediate retrieval or tool-use steps, for accumulating supporting e

safetyarxiv-cs-cl
11 May 2026
Safety

Causal EpiNets: Precision-corrected Bounds on Individual Treatment Effects using Epistemic Neural Networks

DGX agent

arXiv:2605.07065v1 Announce Type: cross Abstract: Individual treatment effects are not point-identified from data. The Probability of Necessity and Sufficiency (PNS) circumvents this limitation by cha

safetyarxiv-cs-ai
11 May 2026
Safety

Cognitive Agent Compilation for Explicit Problem Solver Modeling

DGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

safetyarxiv-cs-ai
11 May 2026
Safety

Conditional generation of antibody sequences with classifier-guided germline-absorbing discrete diffusion

DGX agent

arXiv:2605.06720v1 Announce Type: cross Abstract: Antibody therapeutics are among the most successful modern medicines, yet computationally designing antibodies with desirable binding and developabili

safetyarxiv-cs-ai
11 May 2026
Safety

Confidence-Aware Alignment Makes Reasoning LLMs More Reliable

DGX agent

arXiv:2605.07353v1 Announce Type: new Abstract: Large reasoning models often reach correct answers through flawed intermediate steps, creating a gap between final accuracy and reasoning reliability. E

safetyarxiv-cs-ai
11 May 2026
Safety

Conformal-Style Quantile Analyses for Stochastic Bandits

DGX agent

arXiv:2605.07115v1 Announce Type: new Abstract: Stochastic bandit algorithms are usually analyzed under a mean-reward criterion, yet many problems favor arms with strong upper-tail performance, which

safetyarxiv-cs-lg
11 May 2026
Safety

Contact-Grounded Policy: Dexterous Visuotactile Policy with Generative Contact Grounding

DGX agent

arXiv:2603.05687v3 Announce Type: replace Abstract: Contact-rich dexterous manipulation with multi-finger hands remains an open challenge in robotics because task success depends on multi-point contac

safetyarxiv-cs-ro
11 May 2026
Safety

Cost-Ordered Feasibility for Multi-Armed Bandits with Cost Subsidy

DGX agent

arXiv:2605.07171v1 Announce Type: new Abstract: The classic multi-armed bandit (MAB) problem tackles the challenge of accruing maximum reward while making decisions under uncertainty. However, in appl

safetyarxiv-cs-lg
11 May 2026
Safety

Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences

DGX agent

arXiv:2605.07724v1 Announce Type: cross Abstract: Recursive retraining of generative models poses a critical representation challenge: when synthetic outputs are curated based on a fixed reward signal

safetyarxiv-cs-ai
11 May 2026
Safety

DCGL: Dual-Channel Graph Learning with Large Language Models for Knowledge-Aware Recommendation

DGX agent

arXiv:2605.07314v1 Announce Type: cross Abstract: Knowledge Graphs (KGs) have proven highly effective for recommendation systems by capturing latent item relationships, while recent integration of Lar

safetyarxiv-cs-ai
11 May 2026
Safety

Decentralized Time-Varying Optimization for Streaming Data via Temporal Weighting

DGX agent

arXiv:2605.06971v1 Announce Type: cross Abstract: Classical optimization theory largely focuses on fixed objective functions, whereas many modern learning systems operate in dynamic environments where

safetyarxiv-cs-ai
11 May 2026
Safety

Decoupling Semantics and Fingerprints: A Universal Representation for AI-Generated Image Detection

DGX agent

arXiv:2605.07074v1 Announce Type: new Abstract: Detecting AI-generated images across unseen architectures remains challenging, as existing models often overfit to generator-specific fingerprints and s

safetyarxiv-cs-cv
11 May 2026
Safety

DiffeoMorph: Learning to Morph 3D Shapes Using Differentiable Agent-Based Simulations

DGX agent

arXiv:2512.17129v2 Announce Type: replace Abstract: Biological systems can form complex three-dimensional structures through the collective behavior of agents that share a common update rule and opera

safetyarxiv-cs-lg
11 May 2026
Safety

Differentially Private Auditing Under Strategic Response

DGX agent

arXiv:2605.07674v1 Announce Type: cross Abstract: Regulatory audits of AI systems increasingly rely on differential privacy (DP) to protect training data and model internals. We study audit design whe

safetyarxiv-cs-lg
11 May 2026
Safety

Diffusion-APO: Trajectory-Aware Direct Preference Alignment for Video Diffusion Transformers

DGX agent

arXiv:2605.07503v1 Announce Type: new Abstract: Efficiently aligning large-scale video diffusion models with human intent requires a scalable and trajectory-aware pathway that bridges the inherent dis

safetyarxiv-cs-cv
11 May 2026
Safety

Direction-Flipped Influence Audits Reveal Hidden Structure in Moral Choices of LLMs

DGX agent

arXiv:2602.22831v2 Announce Type: replace-cross Abstract: Moral benchmarks for LLMs typically score models on context-free prompts, implicitly treating the measured choice rate as stable. We test this

safetyarxiv-cs-ai
11 May 2026
Safety

Discovering Multiagent Learning Algorithms with Large Language Models

DGX agent

arXiv:2602.16928v3 Announce Type: replace-cross Abstract: Much of the advancement in Multi-Agent Reinforcement Learning (MARL) for imperfect-information games has historically depended on the manual,

safetyarxiv-cs-ai
11 May 2026
Safety

Does Your Neural Network Extrapolate? Feature Engineering as Identifiability Bias for OOD Generalization

DGX agent

arXiv:2605.07483v1 Announce Type: cross Abstract: Successful deep neural networks discover salient features of data. We show when and why they fail to learn out-of-distribution (OOD)-relevant represen

safetyarxiv-cs-ai
11 May 2026
← Previous
1…192193194195196…257
Next →