AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

GRPO is Secretly a Process Reward Model

DGX agent

arXiv:2509.21154v4 Announce Type: replace-cross Abstract: Process reward models (PRMs) allow for fine-grained credit assignment in reinforcement learning (RL), and seemingly contrast with outcome rewa

safetyarxiv-cs-ai
29 May 2026
Safety

GRUFF: LLM Pronoun Fidelity, Reasoning, and Biases in German

DGX agent

arXiv:2605.30214v1 Announce Type: new Abstract: Third-person singular pronouns have long been used to study stereotypical biases in language models and to test their abilities to reason about referenc

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
safetyarxiv-cs-cl
29 May 2026
Safety

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization

DGX agent

arXiv:2605.29198v1 Announce Type: new Abstract: Group-advantage-based reinforcement learning methods, such as GRPO and DAPO, have demonstrated strong performance across diverse domains, including math

safetyarxiv-cs-cv
29 May 2026
Safety

Harnessing non-adversarial robustness in large language models

DGX agent

arXiv:2605.29816v1 Announce Type: new Abstract: The work presents an approach for addressing the challenge of robustness in Large Language Models (LLMs) to alterations and potential errors caused by s

safetyarxiv-cs-ai
29 May 2026
Safety

How's it going? Reinforcement learning in language models recruits a functional welfare axis

DGX agent

arXiv:2605.30232v1 Announce Type: cross Abstract: How does reinforcement learning shape a language model's internal representations? We present evidence that RL recruits a pre-existing representation

safetyarxiv-cs-cl
29 May 2026
Safety

HPO: Hysteretic Policy Optimization for Stable and Efficient Training under Sparse-Reward Regime

DGX agent

arXiv:2605.30201v1 Announce Type: cross Abstract: We investigate a narrow but common failure mode of GRPO-style reinforcement learning in the context of sparse verifiable rewards: early updates contai

safetyarxiv-cs-ai
29 May 2026
Safety

i have strong reason to believe this is cope. we may find out soon…

DGX agent

i have strong reason to believe this is cope. we may find out soon… Im calling BS on this story. 1. That would be 100,000 employees spending 5k/mo each or 10,000 employees averaging 50k/mo each. No wa

safetygary-marcus--x
29 May 2026
Safety

Improving CLIP Adaptation by Breaking Tail Alignment for Source-Free Cross-Domain Few-Shot Learning

DGX agent

arXiv:2605.29776v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP demonstrate strong zero-shot generalization, but their performance significantly degrades in cross-domain sce

safetyarxiv-cs-cv
29 May 2026
Safety

In-Context Reward Adaptation for Robust Preference Modeling

DGX agent

arXiv:2605.30323v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) typically relies on static reward models to align Large Language Models with human preferences. Howe

safetyarxiv-cs-ai
29 May 2026
Safety

Inferring Code Correctness from Specification

DGX agent

arXiv:2605.29822v1 Announce Type: cross Abstract: Large language models (LLMs) have become integral to modern software development, enabling automated code generation at scale. However, validating the

safetyarxiv-cs-ai
29 May 2026
Safety

Information-Directed Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2605.29405v1 Announce Type: new Abstract: Decision-making from offline datasets typically warm-starts a policy or score model from fixed offline data and then refines it with limited online inte

safetyarxiv-cs-lg
29 May 2026
Safety

It might be time to fill one of these out again. Please remit to myself or @GaryMarcus Thank you.

DGX agent

Gary Marcus is requesting that someone complete a form or document and submit it to him or another person (possibly Gary Marcus himself based on the mention of @GaryMarcus). The post appears to be a r

safetygary-marcus--x
29 May 2026
Safety

It’s a good day when the Pope vouches for your recent comment in Nature.

DGX agent

It’s a good day when the Pope vouches for your recent comment in Nature. The Pope is making exactly our point. LLMs “may imitate or even simulate, but they do not understand.” This is the core epistem

safetygary-marcus--x
29 May 2026
Safety

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too m…

DGX agent

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too much time in/around the big AI Labs (research for the book he

safetygary-marcus--x
29 May 2026
Safety

KGEdit: Ambiguity-Aware Knowledge Graphs for Training-Free Precise Video Generation and Editing

DGX agent

arXiv:2605.29509v1 Announce Type: new Abstract: In recent years, training-free video generation has progressed remarkably. However, when handling complex textual instructions, existing methods still s

safetyarxiv-cs-cv
29 May 2026
Safety

Learning A Simulation-based Visual Policy for Real-world Peg In Unseen Holes

DGX agent

arXiv:2205.04297v2 Announce Type: replace-cross Abstract: This paper proposes a learning-based visual peg-in-hole that enables training with several shapes in simulation, and adapting to arbitrary uns

safetyarxiv-cs-ai
29 May 2026
Safety

Learning to Choose: An Empowerment-Guided Multi-Agent System with semantic communication for Adaptive Method Selection

DGX agent

arXiv:2605.30042v1 Announce Type: new Abstract: Automating scientific computing workflows requires more than generating executable code: autonomous systems must also select appropriate computational s

safetyarxiv-cs-ai
29 May 2026
Safety

Lee Kuan Yew abolished trial by jury in Singapore after determining that it was too easy for defence lawyers to appeal to racial and religio…

DGX agent

Lee Kuan Yew abolished trial by jury in Singapore after determining that it was too easy for defence lawyers to appeal to racial and religious biases of juries in multicultural Singapore. He writes in

safetyelon-musk--x
29 May 2026
Safety

Less Is More: Elevating RAG via Performance-Driven Context Compression

DGX agent

arXiv:2508.19282v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for improving the timeliness of knowledge updates and the factual acc

safetyarxiv-cs-ai
29 May 2026
Safety

LLM-Guided Future Hypotheses for Horizon-Aware Exploration in Multi-Step Robot Manipulation

DGX agent

arXiv:2605.29864v1 Announce Type: new Abstract: Multi-step robot manipulation requires acting under uncertainty about how the scene will evolve, making exploration and policy adaptation challenging. W

safetyarxiv-cs-ro
29 May 2026
Safety

LoMo: Local Modality Substitution for Deeper Vision-Language Fusion

DGX agent

arXiv:2605.30265v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have achieved substantial progress across a wide range of understanding and reasoning tasks, driven by large-scale image

safetyarxiv-cs-cl
29 May 2026
Safety

Looking Beyond Text: Reducing Language bias in Large Vision-Language Models via Multimodal Dual-Attention and Soft-Image Guidance

DGX agent

arXiv:2411.14279v2 Announce Type: replace-cross Abstract: Large vision-language models (LVLMs) have achieved impressive results in various vision-language tasks. However, despite showing promising per

safetyarxiv-cs-cl
29 May 2026
Safety

MARS Policy: Multimodality Only When It Matters

DGX agent

arXiv:2605.29766v1 Announce Type: new Abstract: Imitation learning has become a cornerstone for solving complex robotic manipulation tasks. In particular, multimodality, which enables robots to captur

safetyarxiv-cs-ro
29 May 2026
Safety

Mask the Target: A Plug-and-Play Regularizer Against LoRA Forgetting

DGX agent

arXiv:2605.29498v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become one of the most widely used fine-tuning mechanisms for adapting large language models to new domains, tasks, and u

safetyarxiv-cs-cl
29 May 2026
Safety

MATANet: A Multi-context Attention and Taxonomy-Aware Network for Fine-Grained Underwater Recognition of Marine Species

DGX agent

arXiv:2601.03729v2 Announce Type: replace Abstract: Fine-grained recognition of marine organisms is important for ecological research, biodiversity monitoring, habitat conservation, and evidence-based

safetyarxiv-cs-cv
29 May 2026
Safety

Mean-Field Diffuser: Scaling Offline MARL to Thousands of Agents

DGX agent

arXiv:2605.30190v1 Announce Type: new Abstract: Diffusion-based planning has achieved strong results in single-agent offline reinforcement learning, yet scaling to many-agent systems remains intractab

safetyarxiv-cs-lg
29 May 2026
Safety

MetaRanker: Human-in-the-loop Active Ranking for Metalens Image Quality

DGX agent

arXiv:2605.29212v1 Announce Type: new Abstract: Image quality in modern imaging systems emerges from the coupled effects of the sensor, optics, and computational reconstruction. Ultra-thin metalenses

safetyarxiv-cs-cv
29 May 2026
Safety

MIC: Maximizing Informational Capacity in Adaptive Representations via Isotropic Subspace Alignment

DGX agent

arXiv:2605.29987v1 Announce Type: cross Abstract: Although multi-scales representation learning enables elastic-dimension embeddings, nested subspaces often suffer from dimensional redundancy and spec

safetyarxiv-cs-cl
29 May 2026
Safety

Mining or Synthesis? Rethinking Exploration Efficiency in Iterative Alignment of Mathematical Reasoning

DGX agent

arXiv:2602.05370v3 Announce Type: replace Abstract: Iterative Direct Preference Optimization (DPO) has emerged as a widely used paradigm for aligning Large Language Models on reasoning tasks. Existing

safetyarxiv-cs-cl
29 May 2026
Safety

Mitigating State Aliasing in Vision-Language-Action Models via Inverse Dynamics Learning

DGX agent

arXiv:2605.29577v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a promising framework that unifies perception, reasoning, and control for robot manipulation by adap

safetyarxiv-cs-cv
29 May 2026
Safety

Modality Alignment across Trees on Heterogeneous Hyperbolic Manifolds

DGX agent

arXiv:2510.27391v2 Announce Type: replace Abstract: Modality alignment is critical for vision-language models (VLMs) to effectively integrate information across modalities. However, existing methods e

safetyarxiv-cs-cv
29 May 2026
Safety

Model Fusion via Retrofitting

DGX agent

arXiv:2507.00037v2 Announce Type: replace-cross Abstract: Model fusion seeks to combine independently trained neural networks into a single model without retraining, but is complicated by representati

safetyarxiv-cs-ai
29 May 2026
Safety

Modeling Hierarchical Thinking in Large Reasoning Models

DGX agent

arXiv:2510.22437v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reas

safetyarxiv-cs-ai
29 May 2026
Safety

Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance

DGX agent

arXiv:2605.30187v1 Announce Type: new Abstract: The widespread adoption of AI chatbots in education will drastically change learning, making responsible deployment a critical concern. While large lang

safetyarxiv-cs-ai
29 May 2026
Safety

MonoPhysics: Estimating Geometry, Appearance, and Physical Parameters from Monocular Videos

DGX agent

arXiv:2605.30320v1 Announce Type: new Abstract: Existing inverse physics methods recover physical parameters from multi-view videos, where geometric constraints across views resolve scale and 3D struc

safetyarxiv-cs-cv
29 May 2026
Safety

Multiple executives tell me they’re looking to decrease their AI expenses just as three major AI IPOs are on the horizon. My latest:

DGX agent

Multiple executives tell me they’re looking to decrease their AI expenses just as three major AI IPOs are on the horizon. My latest: CEOs are bargain hunting for AI https://www.axios.com/2026/05/29/ce

safetygary-marcus--x
29 May 2026
Safety

Native Audio-Visual Alignment for Generation

DGX agent

arXiv:2605.30073v1 Announce Type: new Abstract: Joint audio-video generation aims to synthesize temporally synchronized and semantically coherent visual-acoustic content. However, existing open-source

safetyarxiv-cs-cv
29 May 2026
Safety

Neural Operator-Based Surrogate Model for CFD:Helical Coil Steam Generator in Small Modular Reactor

DGX agent

arXiv:2605.30277v1 Announce Type: new Abstract: Real-time thermal-hydraulic simulation is essential for digital twin (DT) technology that supports the safe and efficient operation of small modular rea

safetyarxiv-cs-lg
29 May 2026
Safety

Nobody knows for sure where employment is going and over what time frame. But one thing I can tell you for sure is that the big AI CEO’s hav…

DGX agent

Nobody knows for sure where employment is going and over what time frame. But one thing I can tell you for sure is that the big AI CEO’s have started lying about it. When they tell you know “we are ju

safetygary-marcus--x
29 May 2026
Safety

Note to all staff: Turns out that super tasty AI candy we've been putting out in large bowls isn't free and it isn't necessarily leading to …

DGX agent

Note to all staff: Turns out that super tasty AI candy we've been putting out in large bowls isn't free and it isn't necessarily leading to better products for our customers. From now on, all employee

safetygary-marcus--x
29 May 2026
Safety

Offline Multi-agent Reinforcement Learning via Sequential Score Decomposition

DGX agent

arXiv:2505.05968v3 Announce Type: replace Abstract: Offline cooperative multi-agent reinforcement learning (MARL) faces unique challenges due to distributional shifts, particularly stemming from the h

safetyarxiv-cs-lg
29 May 2026
Safety

Offline Reinforcement Learning with Generative Trajectory Policies

DGX agent

arXiv:2510.11499v2 Announce Type: replace-cross Abstract: Generative models have emerged as a powerful class of policies for offline reinforcement learning (RL) due to their ability to capture complex

safetyarxiv-cs-ai
29 May 2026
Safety

On the Geometry of Games and their Solvers

DGX agent

arXiv:2605.29919v1 Announce Type: new Abstract: A central challenge in game theory and learning systems such as GANs is understanding which algorithms can efficiently compute equilibria across the het

safetyarxiv-cs-ai
29 May 2026
Safety

Online Fair Division with Additional Information

DGX agent

arXiv:2505.24503v3 Announce Type: replace-cross Abstract: We study the problem of fairly allocating indivisible goods to agents in an online setting, where goods arrive sequentially and must be alloca

safetyarxiv-cs-ai
29 May 2026
Safety

Open Problem: Separating Geometric and Algorithmic Compression via Cayley-Table Completion

DGX agent

arXiv:2605.29885v1 Announce Type: new Abstract: Modern statistical learning theory and deep learning characterize generalization primarily in terms of continuous capacity control (e.g., norm-based reg

safetyarxiv-cs-lg
29 May 2026
Safety

Opus 4.8 is insane, nothing will be the same after this model 💀

DGX agent

Gary Marcus expresses strong enthusiasm about Opus 4.8, suggesting it represents a significant breakthrough in AI capabilities. The post implies the model introduces substantial improvements or novel

safetygary-marcus--x
29 May 2026
Safety

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

DGX agent

arXiv:2605.29390v1 Announce Type: new Abstract: Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object

safetyarxiv-cs-cv
29 May 2026
Safety

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

DGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

safetyarxiv-cs-ai
29 May 2026
← Previous
1…172173174175176…302
Next →