AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
All
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,809 results
Safety

Decomposing the Generalization Gap in PROTAC Activity Prediction: Variance Attribution and the Inter-Laboratory Ceiling

DGX agent

arXiv:2605.11764v1 Announce Type: new Abstract: Machine-learning predictors of biochemical activity often exhibit large random-split-to-leave-one-target-out generalisation gaps that have been document

safetyarxiv-cs-lg
13 May 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Diffusion-State Policy Optimization for Masked Diffusion Language Models

DGX agent

arXiv:2602.06462v3 Announce Type: replace Abstract: Masked diffusion language models generate text through iterative masked-token filling, but terminal-only rewards on final completions provide coarse

safetyarxiv-cs-cl
13 May 2026
Safety

Discrete Flow Matching for Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2605.12379v1 Announce Type: new Abstract: Many reinforcement learning (RL) tasks have discrete action spaces, but most generative policy methods based on diffusion and flow matching are designed

safetyarxiv-cs-lg
13 May 2026
Safety

Dissecting Discrete Soft Actor-Critic: Limitations and Principled Alternatives

DGX agent

arXiv:2509.09838v2 Announce Type: replace Abstract: While Soft Actor-Critic (SAC) is highly effective in continuous control, its discrete counterpart (DSAC) performs poorly on challenging discrete-act

safetyarxiv-cs-lg
13 May 2026
Safety

Do multi-agent systems make LLM reasoning better? Most AI devs assume that it should. But this new paper shows that this is often not the ca…

DGX agent

Do multi-agent systems make LLM reasoning better? Most AI devs assume that it should. But this new paper shows that this is often not the case. It ran 22,500 deterministic trajectories across GAIA, SW

safetydair-ai--x
13 May 2026
Safety

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion

DGX agent

arXiv:2505.18780v3 Announce Type: replace-cross Abstract: Achieving versatile humanoid locomotion with a single policy presents a critical scalability challenge. Prevailing methods often rely on disti

safetyarxiv-cs-lg
13 May 2026
Safety

ECTO: Exogenous-Conditioned Temporal Operator for Ultra-Short-Term Wind Power Forecasting

DGX agent

arXiv:2605.12196v1 Announce Type: new Abstract: Accurate ultra-short-term wind power forecasting is critical for grid dispatch and reserve management, yet remains challenging due to the non-stationary

safetyarxiv-cs-lg
13 May 2026
Safety

EHR-RAGp: Retrieval-Augmented Prototype-Guided Foundation Model for Electronic Health Records

DGX agent

arXiv:2605.12335v1 Announce Type: cross Abstract: Electronic Health Records (EHR) contain rich longitudinal patient information and are widely used in predictive modeling applications. However, effect

safetyarxiv-cs-lg
13 May 2026
Safety

Emergent Communication between Heterogeneous Visual Agents through Decentralized Learning

DGX agent

arXiv:2605.11695v1 Announce Type: new Abstract: Symbols are shared, but perception is private. We study emergent communication between heterogeneous visual agents through decentralized learning, askin

safetyarxiv-cs-cv
13 May 2026
Safety

Enabling clinical use of foundation models for computational pathology

DGX agent

arXiv:2602.22347v2 Announce Type: replace Abstract: Foundation models for computational pathology are expected to facilitate the development of high-performing, generalisable deep learning systems. Ho

safetyarxiv-cs-cv
13 May 2026
Safety

Enabling Performant and Flexible Model-Internal Observability for LLM Inference

DGX agent

arXiv:2605.11093v1 Announce Type: new Abstract: Today's inference-time workloads increasingly depend on timely access to a model's internal states. We present DMI-Lib, a high-speed deep model inspecto

safetyarxiv-cs-lg
13 May 2026
Safety

Enforcing Constraints in Generative Sampling via Adaptive Correction Scheduling

DGX agent

arXiv:2605.11214v1 Announce Type: new Abstract: Hard constraints in generative sampling are typically enforced by projection, applied either once at the end of sampling or after every update. This bin

safetyarxiv-cs-lg
13 May 2026
Safety

Enhancing Multilingual Counterfactual Generation through Alignment-as-Preference Optimization

DGX agent

arXiv:2605.11632v1 Announce Type: new Abstract: Self-generated counterfactual explanations (SCEs) are minimally modified inputs (minimality) generated by large language models (LLMs) that flip their o

safetyarxiv-cs-cl
13 May 2026
Safety

Enhancing Target-Guided Proactive Dialogue Systems via Conversational Scenario Modeling and Intent-Keyword Bridging

DGX agent

arXiv:2605.11964v1 Announce Type: new Abstract: A target-guided proactive dialogue system aims to steer conversations proactively toward pre-defined targets, such as designated keywords or specific to

safetyarxiv-cs-cl
13 May 2026
Safety

Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control

DGX agent

arXiv:2605.11775v1 Announce Type: cross Abstract: Policy entropy has emerged as a fundamental measure for understanding and controlling exploration in reinforcement learning with verifiable rewards (R

safetyarxiv-cs-cl
13 May 2026
Safety

Epistemic Uncertainty for Test-Time Discovery

DGX agent

arXiv:2605.11328v1 Announce Type: new Abstract: Automated scientific discovery using large language models relies on identifying genuinely novel solutions. Standard reinforcement learning penalizes hi

safetyarxiv-cs-lg
13 May 2026
Safety

Events as Triggers for Behavioral Diversity in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.12388v1 Announce Type: cross Abstract: Effective multi-agent cooperation requires agents to adopt diverse behaviors as task conditions evolve-and to do so at the right moment. Yet, current

safetyarxiv-cs-lg
13 May 2026
Safety

EvoNav: Evolutionary Reward Function Design for Robot Navigation with Large Language Models

DGX agent

arXiv:2605.11859v1 Announce Type: new Abstract: Robot navigation is a crucial task with applications to social robots in dynamic human environments. While Reinforcement Learning (RL) has shown great p

safetyarxiv-cs-ro
13 May 2026
Safety

Expected Batch Optimal Transport Plans and Consequences for Flow Matching

DGX agent

arXiv:2605.12174v1 Announce Type: new Abstract: Solving optimal transport (OT) on random minibatches is a common surrogate for exact OT in large-scale learning. In flow matching (FM), this surrogate i

safetyarxiv-cs-lg
13 May 2026
Safety

Fair Conformal Classification via Learning Representation-Based Groups

DGX agent

arXiv:2605.12195v1 Announce Type: new Abstract: Conformal prediction methods provide statistically rigorous marginal coverage guarantees for machine learning models, but such guarantees fail to accoun

safetyarxiv-cs-lg
13 May 2026
Safety

Fed-BAC: Federated Bandit-Guided Additive Clustering in Hierarchical Federated Learning

DGX agent

arXiv:2605.11815v1 Announce Type: new Abstract: Hierarchical federated learning (HFL) leverages edge servers for partial aggregation in edge computing. Yet existing FL methods lack mechanisms for join

safetyarxiv-cs-lg
13 May 2026
Safety

FedOUI: OUI-Guided Client Weighting for Federated Aggregation

DGX agent

arXiv:2605.11571v1 Announce Type: new Abstract: Federated learning usually aggregates client updates using dataset size or gradient-level criteria, while overlooking internal signals about how each cl

safetyarxiv-cs-lg
13 May 2026
Safety

FedSurrogate: Backdoor Defense in Federated Learning via Layer Criticality and Surrogate Replacement

DGX agent

arXiv:2605.11122v1 Announce Type: cross Abstract: Federated Learning remains highly susceptible to backdoor attacks--malicious clients inject targeted behaviours into the global model. Existing defens

safetyarxiv-cs-lg
13 May 2026
Safety

Few-Shot Synthetic Data Generation with Diffusion Models for Downstream Vision Tasks

DGX agent

arXiv:2605.11898v1 Announce Type: new Abstract: Class imbalance is a persistent challenge in visual recognition, particularly in safety-critical domains where collecting positive examples is expensive

safetyarxiv-cs-cv
13 May 2026
Safety

Fill the GAP: A Granular Alignment Paradigm for Visual Reasoning in Multimodal Large Language Models

DGX agent

arXiv:2605.12374v1 Announce Type: new Abstract: Visual latent reasoning lets a multimodal large language model (MLLM) create intermediate visual evidence as continuous tokens, avoiding external tools

safetyarxiv-cs-cv
13 May 2026
Safety

ForceFlow: Learning to Feel and Act via Contact-Driven Flow Matching

DGX agent

arXiv:2605.11048v1 Announce Type: new Abstract: Existing imitation learning methods enable robots to interact autonomously with the physical environment. However, contact-rich manipulation tasks remai

safetyarxiv-cs-ro
13 May 2026
Safety

From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation

DGX agent

arXiv:2605.11613v1 Announce Type: new Abstract: On-policy self-distillation has emerged as a promising paradigm for post-training language models, in which the model conditions on environment feedback

safetyarxiv-cs-lg
13 May 2026
Safety

From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation

DGX agent

arXiv:2605.12167v1 Announce Type: cross Abstract: Video generation models offer a promising imagination mechanism for robot manipulation by predicting long-horizon future observations, but effectively

safetyarxiv-cs-cv
13 May 2026
Safety

From Message-Passing to Linearized Graph Sequence Models

DGX agent

arXiv:2605.12358v1 Announce Type: new Abstract: Message-passing based approaches form the default backbone of most learning architectures on graph-structured data. However, the rapid progress of moder

safetyarxiv-cs-lg
13 May 2026
Safety

Fused Gromov-Wasserstein Distance with Feature Selection

DGX agent

arXiv:2605.12161v1 Announce Type: new Abstract: Fused Gromov-Wasserstein (FGW) distances provide a principled framework for comparing objects by jointly aligning structure and node features. However,

safetyarxiv-cs-lg
13 May 2026
Safety

GEAR: Granularity-Adaptive Advantage Reweighting for LLM Agents via Self-Distillation

DGX agent

arXiv:2605.11853v1 Announce Type: cross Abstract: Reinforcement learning has become a widely used post-training approach for LLM agents, where training commonly relies on outcome-level rewards that pr

safetyarxiv-cs-cl
13 May 2026
Safety

Generative AI for Visualizing Highway Construction Hazards Through Synthetic Images and Temporal Sequences

DGX agent

arXiv:2605.11276v1 Announce Type: new Abstract: Highway construction workers face a high risk of serious injury or death. Image-based training materials depicting hazardous scenarios are essential for

safetyarxiv-cs-cv
13 May 2026
Safety

Generative AI has not made the world a better place.

DGX agent

Generative AI has not made the world a better place. Since 1893, Princeton professors have left the room when students take their final exams. The idea was that if you treat students honorably, they w

safetygary-marcus--x
13 May 2026
Safety

Generative climate downscaling enables high-resolution compound risk assessment by preserving multivariate dependencies

DGX agent

arXiv:2605.11531v1 Announce Type: cross Abstract: Physics-based climate projections using general circulation models are essential for assessing future risks, but their coarse resolution limits region

safetyarxiv-cs-lg
13 May 2026
Safety

Gradient-Free Noise Optimization for Reward Alignment in Generative Models

DGX agent

arXiv:2605.11347v1 Announce Type: cross Abstract: Existing reward alignment methods for diffusion and flow models rely on multi-step stochastic trajectories, making them difficult to extend to determi

safetyarxiv-cs-cv
13 May 2026
Safety

GRAFT: Graph-Tokenized LLMs for Tool Planning

DGX agent

arXiv:2605.11706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to complete complex tasks by selecting and coordinating external tools across multiple steps. This re

safetyarxiv-cs-lg
13 May 2026
Safety

GRASP: Guided Residual Adapters with Sample-wise Partitioning

DGX agent

arXiv:2512.01675v2 Announce Type: replace Abstract: Text-to-image flow matching transformers degrade sharply in long-tail settings: tail-class outputs collapse in fidelity and diversity, limiting thei

safetyarxiv-cs-cv
13 May 2026
Safety

Hey @Elonmusk I laid out the core of your lawyer’s case against Altman’s credibility almost three years ago in this tweet. Aged well!

DGX agent

Hey @Elonmusk I laid out the core of your lawyer’s case against Altman’s credibility almost three years ago in this tweet. Aged well! Remember how Sam Altman told the US Senate he had no “direct” inve

safetygary-marcus--x
13 May 2026
Safety

Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers

DGX agent

arXiv:2605.11556v1 Announce Type: cross Abstract: Solving complex long-horizon tasks requires strong planning and reasoning capabilities. Although datasets with explicit chain-of-thought (CoT) rationa

safetyarxiv-cs-lg
13 May 2026
Safety

How Does Differential Privacy Affect Social Bias in LLMs? A Systematic Evaluation

DGX agent

arXiv:2605.11195v1 Announce Type: new Abstract: Large language models (LLMs) trained on web-scale corpora can memorize sensitive training data, posing significant privacy risks. Differential privacy (

safetyarxiv-cs-cl
13 May 2026
Safety

How far can bias go? Tracing bias from pretraining data to alignment

DGX agent

arXiv:2411.19240v2 Announce Type: replace Abstract: As LLMs are increasingly integrated into user-facing applications, addressing biases that perpetuate societal inequalities is crucial. While much wo

safetyarxiv-cs-cl
13 May 2026
Safety

'I applied to be pope': Losing grip on reality while using ChatGPT. AFP spoke to members of a support group for people suffering from AI-ind…

DGX agent

'I applied to be pope': Losing grip on reality while using ChatGPT. AFP spoke to members of a support group for people suffering from AI-induced delusion or psychosis. All warned that the world has to

safetygary-marcus--x
13 May 2026
Safety

i just can’t get over the first sentence here; the absolute confidence with which is it said, and the high likelihood that is epically wrong…

DGX agent

i just can’t get over the first sentence here; the absolute confidence with which is it said, and the high likelihood that is epically wrong. hereby nominating it to the list of sentences that within

safetygary-marcus--x
13 May 2026
Safety

Images in Sentences: Scaling Interleaved Instructions for Unified Visual Generation

DGX agent

arXiv:2605.12305v1 Announce Type: new Abstract: While recent advancements in multimodal language models have enabled image generation from expressive multi-image instructions, existing methods struggl

safetyarxiv-cs-cv
13 May 2026
Safety

In-Context Multi-Objective Optimization

DGX agent

arXiv:2512.11114v2 Announce Type: replace Abstract: Balancing competing objectives is omnipresent across disciplines, from drug design to autonomous systems. Multi-objective Bayesian optimization is a

safetyarxiv-cs-lg
13 May 2026
Safety

Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning

DGX agent

arXiv:2605.11889v1 Announce Type: new Abstract: Collaborative machine learning involves training high-quality models using datasets from a number of sources. To incentivize sources to share data, exis

safetyarxiv-cs-lg
13 May 2026
Safety

Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning

DGX agent

arXiv:2605.11235v1 Announce Type: new Abstract: In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment

safetyarxiv-cs-lg
13 May 2026
Safety

Interpreting Context-Aware Human Preferences for Multi-Objective Robot Navigation

DGX agent

arXiv:2603.17510v2 Announce Type: replace Abstract: Robots operating in human-shared environments must not only achieve task-level navigation objectives such as safety and efficiency, but also adapt t

safetyarxiv-cs-ro
13 May 2026
← Previous
1…182183184185186…267
Next →