AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
Safety

reply to Hinton’s reply to me, for additional context:

DGX agent

reply to Hinton’s reply to me, for additional context: Dear @geoffreyhinton, I literally never said that AI systems “JUST regurgitate”; that’s plainly false. I don’t believe it, and I didn’t say it. (

safetygary-marcus--x
11 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Resource-Element Energy Difference for Noncoherent Over-the-Air Federated Learning

DGX agent

arXiv:2605.07263v1 Announce Type: cross Abstract: Over-the-air federated learning (OTA-FL) reduces uplink latency by exploiting waveform superposition, but conventional analog aggregation schemes typi

safetyarxiv-cs-ai
11 May 2026
Safety

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding

DGX agent

arXiv:2605.07575v1 Announce Type: cross Abstract: Proactive streaming video understanding requires Video-LLMs to decide when to respond as a video unfolds, a task where existing methods often fall sho

safetyarxiv-cs-ai
11 May 2026
Safety

Response Time Enhances Alignment with Heterogeneous Preferences

DGX agent

arXiv:2605.06987v1 Announce Type: new Abstract: Aligning large language models (LLMs) to human preferences typically relies on aggregating pooled feedback into a single reward model. However, this sta

safetyarxiv-cs-lg
11 May 2026
Safety

Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

DGX agent

arXiv:2605.07331v1 Announce Type: cross Abstract: Reinforcement learning, including reinforcement learning with verifiable rewards (RLVR), has emerged as a powerful approach for LLM post-training. Cen

safetyarxiv-cs-ai
11 May 2026
Safety

RIDER: 3D RNA Inverse Design with Reinforcement Learning-Guided Diffusion

DGX agent

arXiv:2602.16548v2 Announce Type: replace Abstract: The inverse design of RNA three-dimensional (3D) structures is crucial for engineering functional RNAs in synthetic biology and therapeutics. While

safetyarxiv-cs-lg
11 May 2026
Safety

Risk-Consistent Multiclass Learning from Random Label-Subset Membership Queries

DGX agent

arXiv:2605.07413v1 Announce Type: new Abstract: Obtaining accurate class labels is often costly or unreliable, and may also be limited by privacy or other practical conditions. Compared with asking an

safetyarxiv-cs-lg
11 May 2026
Safety

Robustness of Refugee-Matching Gains to Off-Policy Evaluation Choices

DGX agent

arXiv:2605.06686v1 Announce Type: new Abstract: Previous research has investigated the potential of refugee matching for boosting refugee outcomes, first considered by Bansak et al. (2018). This paper

safetyarxiv-cs-lg
11 May 2026
Safety

Rollback-Free Stable Brick Structures Generation

DGX agent

arXiv:2605.06947v1 Announce Type: new Abstract: While autoregressive models have advanced 3D generation, creating physically stable brick structures remains a challenge due to the strict requirements

safetyarxiv-cs-lg
11 May 2026
Safety

Rubric-based On-policy Distillation

DGX agent

arXiv:2605.07396v1 Announce Type: cross Abstract: On-policy distillation (OPD) is a powerful paradigm for model alignment, yet its reliance on teacher logits restricts its application to white-box sce

safetyarxiv-cs-ai
11 May 2026
Safety

Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence

DGX agent

arXiv:2605.06230v2 Announce Type: replace Abstract: As large models evolve from conversational assistants into autonomous agents, challenges increasingly arise from long-horizon decision making, tool

safetyarxiv-cs-ai
11 May 2026
Safety

SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions

DGX agent

arXiv:2605.07102v1 Announce Type: new Abstract: Evaluating literary quality requires assessing interpretive dimensions such as cultural representation, emotional depth, and philosophical sophisticatio

safetyarxiv-cs-cl
11 May 2026
Safety

Same Signal, Opposite Meaning: Direction-Informed Adaptive Learning for LLM Agents

DGX agent

arXiv:2605.06908v1 Announce Type: cross Abstract: Adaptive test-time compute for LLM agents aims to invoke extra computation only when it improves performance. Existing methods typically use confidenc

safetyarxiv-cs-ai
11 May 2026
Safety

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

DGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

safetyarxiv-cs-cv
11 May 2026
Safety

Science publishing giant Elsevier has joined the dozens of firms and individuals suing artificial intelligence companies over their alleged …

DGX agent

Science publishing giant Elsevier has joined the dozens of firms and individuals suing artificial intelligence companies over their alleged use of copyrighted works in training AI models https://go.na

safetygary-marcus--x
11 May 2026
Safety

Self-Programmed Execution for Language-Model Agents

DGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

safetyarxiv-cs-ai
11 May 2026
Safety

Serious question: Should I write a short book called 7 lies about AI that never die?

DGX agent

Serious question: Should I write a short book called 7 lies about AI that never die? AI hype has become a giant game of bait and switch. the bait: we are going to make an AI that can solve any problem

safetygary-marcus--x
11 May 2026
Safety

SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents

DGX agent

arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa

safetyarxiv-cs-lg
11 May 2026
Safety

SimCT: Recovering Lost Supervision for Cross-Tokenizer On-Policy Distillation

DGX agent

arXiv:2605.07711v1 Announce Type: new Abstract: On-policy distillation (OPD) is a standard tool for transferring teacher behavior to a smaller student, but it implicitly assumes that teacher and stude

safetyarxiv-cs-cl
11 May 2026
Safety

Since Hinton has actually replied let me clarify some things - LLMS don’t *always* regurgitate - LLMs don’t literally store full texts - but…

DGX agent

Since Hinton has actually replied let me clarify some things - LLMS don’t *always* regurgitate - LLMs don’t literally store full texts - but given the mechanisms that they use they do sometimes regurg

safetygary-marcus--x
11 May 2026
Safety

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

DGX agent

arXiv:2605.06130v2 Announce Type: replace Abstract: A persistent skill library allows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupl

safetyarxiv-cs-ai
11 May 2026
Safety

Slowly Annealed Langevin Dynamics: Theory and Applications to Training-Free Guided Generation

DGX agent

arXiv:2605.07950v1 Announce Type: new Abstract: We study Slowly Annealed Langevin Dynamics (SALD), a sampler for tracking a path of moving target distributions and approximating the terminal target th

safetyarxiv-cs-lg
11 May 2026
Safety

SOD: Step-wise On-policy Distillation for Small Language Model Agents

DGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

safetyarxiv-cs-ai
11 May 2026
Safety

Sources: the White House's Office of the National Cyber Director and Commerce Department's CAISI are fighting over which agency should lead AI model evaluations (Washington Post)

DGX agent

Washington Post: Sources: the White House's Office of the National Cyber Director and Commerce Department's CAISI are fighting over which agency should lead AI model evaluations — As the White House g

safetytechmeme
11 May 2026
Safety

SparseRL-Sync: Lossless Weight Synchronization with ~100x Less Communication

DGX agent

arXiv:2605.07330v1 Announce Type: cross Abstract: In large-scale reinforcement learning (RL) systems with decoupled Trainer-Rollout execution, the Trainer must regularly synchronize policy weights to

safetyarxiv-cs-ai
11 May 2026
Safety

Stabilized neural Hamilton--Jacobi--Bellman solvers: Error analysis and applications in model-based reinforcement learning

DGX agent

arXiv:2605.07116v1 Announce Type: cross Abstract: Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is

safetyarxiv-cs-ai
11 May 2026
Safety

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration

DGX agent

arXiv:2512.23927v2 Announce Type: replace-cross Abstract: Fitted Q-iteration (FQI) and soft FQI are widely used value-based methods for offline reinforcement learning, but their standard stability gua

safetyarxiv-cs-lg
11 May 2026
Safety

STDA-Net: Spectrogram-Based Domain Adaptation for cross-dataset Sleep Stage Classification

DGX agent

arXiv:2605.06736v1 Announce Type: cross Abstract: Accurate sleep stage classification across datasets remains challenging due to variability in EEG channel montages, sampling rates, recording environm

safetyarxiv-cs-ai
11 May 2026
Safety

Structured Role-Aware Policy Optimization for Multimodal Reasoning

DGX agent

arXiv:2605.07274v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR), especially with Group Relative Policy Optimization (GRPO), has shown strong potential for improvi

safetyarxiv-cs-ai
11 May 2026
Safety

Supervised sparse auto-encoders for interpretable and compositional representations

DGX agent

arXiv:2602.00924v2 Announce Type: replace Abstract: Sparse auto-encoders (SAEs) have re-emerged as a prominent method for mechanistic interpretability, yet they face two significant challenges: the no

safetyarxiv-cs-ai
11 May 2026
Safety

Temporal Attention for Adaptive Control of Euler-Lagrange Systems with Unobservable Memory

DGX agent

arXiv:2605.06877v1 Announce Type: new Abstract: Adaptive control of Euler-Lagrange systems is challenging when friction is governed by a finite-horizon internal state that is not directly observable f

safetyarxiv-cs-lg
11 May 2026
Safety

Temporal Smoothness Doubly Robust Learning for Debiased Knowledge Tracing

DGX agent

arXiv:2605.05958v2 Announce Type: replace Abstract: Knowledge Tracing (KT) is fundamental to intelligent education systems, yet relies on educational logs that are selectively observed. The non-random

safetyarxiv-cs-ai
11 May 2026
Safety

TextLDM: Language Modeling with Continuous Latent Diffusion

DGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

safetyarxiv-cs-cl
11 May 2026
Safety

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone f…

DGX agent

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone from their website, link no longer found, and I can't get any

safetygary-marcus--x
11 May 2026
Safety

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

DGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

safetyarxiv-cs-ro
11 May 2026
Safety

The Effect of Mini-Batch Noise on the Implicit Bias of Adam

DGX agent

arXiv:2602.01642v2 Announce Type: replace-cross Abstract: With limited high-quality data and growing compute, multi-epoch training is gaining back its importance across sub-areas of deep learning. Ada

safetyarxiv-cs-ai
11 May 2026
Safety

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

DGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

safetyarxiv-cs-ai
11 May 2026
Safety

The Limits of AI-Driven Allocation: Optimal Screening under Aleatoric Uncertainty

DGX agent

arXiv:2605.07979v1 Announce Type: new Abstract: The rise of machine learning has shifted targeted resource allocation in policy and humanitarian settings toward algorithmic targeting based on predicte

safetyarxiv-cs-ai
11 May 2026
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Safety

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

DGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

safetyarxiv-cs-cl
11 May 2026
Safety

This is what a useless hype lifecycle looks like.

DGX agent

Gary Marcus critiques the typical hype cycle pattern where emerging technologies experience inflated expectations followed by inevitable disappointment. The post likely illustrates this cycle using a

safetygary-marcus--x
11 May 2026
Safety

totally worth $10 trillion a year

DGX agent

This post from AI researcher Gary Marcus likely discusses the enormous economic value or potential return on investment related to artificial intelligence developments, suggesting AI's worth or impact

safetygary-marcus--x
11 May 2026
Safety

Toward Better Geometric Representations for Molecule Generative Models

DGX agent

arXiv:2605.07693v1 Announce Type: new Abstract: Geometric representation-conditioned molecule generation provides an effective paradigm that decouples molecule representation modeling from structure g

safetyarxiv-cs-lg
11 May 2026
Safety

Towards Differentially Private Reinforcement Learning with General Function Approximation

DGX agent

arXiv:2605.07049v1 Announce Type: cross Abstract: We present the first theoretical guarantees for differentially private online reinforcement learning (RL) with general function approximation, extendi

safetyarxiv-cs-ai
11 May 2026
Safety

Towards Fairness under Label Bias in Image Segmentation: Impact, Measurement and Mitigation

DGX agent

arXiv:2605.06891v1 Announce Type: new Abstract: Labeled datasets reflect the biases of their annotation pipelines, which sometimes introduce label bias: group-conditional label errors that cause syste

safetyarxiv-cs-cv
11 May 2026
Safety

TRACE: Transport Alignment Conformal Prediction via Diffusion and Flow Matching Models

DGX agent

arXiv:2605.07100v1 Announce Type: cross Abstract: Constructing valid and informative conformal prediction regions for multi-dimensional outputs remains a fundamental challenge. While conformal predict

safetyarxiv-cs-lg
11 May 2026
Safety

Training-Free Multimodal Large Language Model Orchestration

DGX agent

arXiv:2508.10016v3 Announce Type: replace Abstract: Building interactive omni-modal assistants often relies on end-to-end multimodal alignment to fuse heterogeneous modalities, which incurs substantia

safetyarxiv-cs-cl
11 May 2026
Safety

TRAJGANR: Trajectory-Centric Urban Multimodal Learning via Geospatially Aligned Neural Representations

DGX agent

arXiv:2605.06990v1 Announce Type: new Abstract: Multimodal self-supervised learning (MSSL) has emerged as a key paradigm for pretraining geospatial foundation models. However, existing geospatial MSSL

safetyarxiv-cs-cv
11 May 2026
← Previous
1…226227228229230…302
Next →