AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,357 results
11 May 2026

SARA: Semantically Adaptive Relational Alignment for Video Diffusion Models

SafetyDGX agent

arXiv:2605.07800v1 Announce Type: new Abstract: Recent video diffusion models (VDMs) synthesize visually convincing clips, yet still drop entities, mis-bind attributes, and weaken the interactions spe

Science publishing giant Elsevier has joined the dozens of firms and individuals suing artificial intelligence companies over their alleged …

SafetyDGX agent

Science publishing giant Elsevier has joined the dozens of firms and individuals suing artificial intelligence companies over their alleged use of copyrighted works in training AI models https://go.na

Self-Programmed Execution for Language-Model Agents

SafetyDGX agent

arXiv:2605.06898v1 Announce Type: new Abstract: At the heart of existing language model agents is a fixed orchestrator program responsible for the state transition between consecutive turns. This pape

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Serious question: Should I write a short book called 7 lies about AI that never die?

SafetyDGX agent

Serious question: Should I write a short book called 7 lies about AI that never die? AI hype has become a giant game of bait and switch. the bait: we are going to make an AI that can solve any problem

SHARP: A Self-Evolving Human-Auditable Rubric Policy for Financial Trading Agents

SafetyDGX agent

arXiv:2605.06822v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed for autonomous financial trading, a domain requiring continuous adaptation to noisy, non-stationa

SimCT: Recovering Lost Supervision for Cross-Tokenizer On-Policy Distillation

SafetyDGX agent

arXiv:2605.07711v1 Announce Type: new Abstract: On-policy distillation (OPD) is a standard tool for transferring teacher behavior to a smaller student, but it implicitly assumes that teacher and stude

Since Hinton has actually replied let me clarify some things - LLMS don’t *always* regurgitate - LLMs don’t literally store full texts - but…

SafetyDGX agent

Since Hinton has actually replied let me clarify some things - LLMS don’t *always* regurgitate - LLMs don’t literally store full texts - but given the mechanisms that they use they do sometimes regurg

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning

SafetyDGX agent

arXiv:2605.06130v2 Announce Type: replace Abstract: A persistent skill library allows language model agents to reuse successful strategies across tasks. Maintaining such a library requires three coupl

Slowly Annealed Langevin Dynamics: Theory and Applications to Training-Free Guided Generation

SafetyDGX agent

arXiv:2605.07950v1 Announce Type: new Abstract: We study Slowly Annealed Langevin Dynamics (SALD), a sampler for tracking a path of moving target distributions and approximating the terminal target th

SOD: Step-wise On-policy Distillation for Small Language Model Agents

SafetyDGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

Sources: the White House's Office of the National Cyber Director and Commerce Department's CAISI are fighting over which agency should lead AI model evaluations (Washington Post)

SafetyDGX agent

Washington Post: Sources: the White House's Office of the National Cyber Director and Commerce Department's CAISI are fighting over which agency should lead AI model evaluations — As the White House g

SparseRL-Sync: Lossless Weight Synchronization with ~100x Less Communication

SafetyDGX agent

arXiv:2605.07330v1 Announce Type: cross Abstract: In large-scale reinforcement learning (RL) systems with decoupled Trainer-Rollout execution, the Trainer must regularly synchronize policy weights to

Stabilized neural Hamilton--Jacobi--Bellman solvers: Error analysis and applications in model-based reinforcement learning

SafetyDGX agent

arXiv:2605.07116v1 Announce Type: cross Abstract: Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration

SafetyDGX agent

arXiv:2512.23927v2 Announce Type: replace-cross Abstract: Fitted Q-iteration (FQI) and soft FQI are widely used value-based methods for offline reinforcement learning, but their standard stability gua

STDA-Net: Spectrogram-Based Domain Adaptation for cross-dataset Sleep Stage Classification

SafetyDGX agent

arXiv:2605.06736v1 Announce Type: cross Abstract: Accurate sleep stage classification across datasets remains challenging due to variability in EEG channel montages, sampling rates, recording environm

Structured Role-Aware Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2605.07274v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR), especially with Group Relative Policy Optimization (GRPO), has shown strong potential for improvi

Supervised sparse auto-encoders for interpretable and compositional representations

SafetyDGX agent

arXiv:2602.00924v2 Announce Type: replace Abstract: Sparse auto-encoders (SAEs) have re-emerged as a prominent method for mechanistic interpretability, yet they face two significant challenges: the no

Temporal Attention for Adaptive Control of Euler-Lagrange Systems with Unobservable Memory

SafetyDGX agent

arXiv:2605.06877v1 Announce Type: new Abstract: Adaptive control of Euler-Lagrange systems is challenging when friction is governed by a finite-horizon internal state that is not directly observable f

Temporal Smoothness Doubly Robust Learning for Debiased Knowledge Tracing

SafetyDGX agent

arXiv:2605.05958v2 Announce Type: replace Abstract: Knowledge Tracing (KT) is fundamental to intelligent education systems, yet relies on educational logs that are selectively observed. The non-random

TextLDM: Language Modeling with Continuous Latent Diffusion

SafetyDGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone f…

SafetyDGX agent

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone from their website, link no longer found, and I can't get any

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

SafetyDGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

The Effect of Mini-Batch Noise on the Implicit Bias of Adam

SafetyDGX agent

arXiv:2602.01642v2 Announce Type: replace-cross Abstract: With limited high-quality data and growing compute, multi-epoch training is gaining back its importance across sub-areas of deep learning. Ada

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

SafetyDGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

The Limits of AI-Driven Allocation: Optimal Screening under Aleatoric Uncertainty

SafetyDGX agent

arXiv:2605.07979v1 Announce Type: new Abstract: The rise of machine learning has shifted targeted resource allocation in policy and humanitarian settings toward algorithmic targeting based on predicte

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

SafetyDGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

SafetyDGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

This is what a useless hype lifecycle looks like.

SafetyDGX agent

Gary Marcus critiques the typical hype cycle pattern where emerging technologies experience inflated expectations followed by inevitable disappointment. The post likely illustrates this cycle using a

totally worth $10 trillion a year

SafetyDGX agent

This post from AI researcher Gary Marcus likely discusses the enormous economic value or potential return on investment related to artificial intelligence developments, suggesting AI's worth or impact

Toward Better Geometric Representations for Molecule Generative Models

SafetyDGX agent

arXiv:2605.07693v1 Announce Type: new Abstract: Geometric representation-conditioned molecule generation provides an effective paradigm that decouples molecule representation modeling from structure g

Towards Differentially Private Reinforcement Learning with General Function Approximation

SafetyDGX agent

arXiv:2605.07049v1 Announce Type: cross Abstract: We present the first theoretical guarantees for differentially private online reinforcement learning (RL) with general function approximation, extendi

Towards Fairness under Label Bias in Image Segmentation: Impact, Measurement and Mitigation

SafetyDGX agent

arXiv:2605.06891v1 Announce Type: new Abstract: Labeled datasets reflect the biases of their annotation pipelines, which sometimes introduce label bias: group-conditional label errors that cause syste

TRACE: Transport Alignment Conformal Prediction via Diffusion and Flow Matching Models

SafetyDGX agent

arXiv:2605.07100v1 Announce Type: cross Abstract: Constructing valid and informative conformal prediction regions for multi-dimensional outputs remains a fundamental challenge. While conformal predict

Training-Free Multimodal Large Language Model Orchestration

SafetyDGX agent

arXiv:2508.10016v3 Announce Type: replace Abstract: Building interactive omni-modal assistants often relies on end-to-end multimodal alignment to fuse heterogeneous modalities, which incurs substantia

TRAJGANR: Trajectory-Centric Urban Multimodal Learning via Geospatially Aligned Neural Representations

SafetyDGX agent

arXiv:2605.06990v1 Announce Type: new Abstract: Multimodal self-supervised learning (MSSL) has emerged as a key paradigm for pretraining geospatial foundation models. However, existing geospatial MSSL

UFT: Unifying Fine-Tuning of SFT and RLHF/DPO/UNA through a Generalized Implicit Reward Function

SafetyDGX agent

arXiv:2410.21438v3 Announce Type: replace Abstract: By pretraining on trillions of tokens, an LLM gains the capability of text generation. However, to enhance its utility and reduce potential harm, SF

UNA: A Unified Supervised Framework for Efficient LLM Alignment Across Feedback Types

SafetyDGX agent

arXiv:2408.15339v4 Announce Type: replace-cross Abstract: RL alignment methods, including RLHF and DPO, are primarily based on pairwise preference data. Although scalar or score-based feedback has bee

UniD-Shift: Towards Unified Semantic Segmentation via Interpretable Share-Private Multimodal Decomposition

SafetyDGX agent

arXiv:2605.07356v1 Announce Type: new Abstract: Semantic segmentation of large-scale 3D point clouds is crucial for applications such as autonomous driving and urban digital twins. However, the sparse

VDEGaussian: Video Diffusion Enhanced 4D Gaussian Splatting for Dynamic Urban Scenes Modeling

SafetyDGX agent

arXiv:2508.02129v2 Announce Type: replace Abstract: Dynamic urban scene modeling is a rapidly evolving area with broad applications. While current approaches leveraging neural radiance fields or Gauss

VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training

SafetyDGX agent

arXiv:2602.10693v3 Announce Type: replace-cross Abstract: Off-policy updates are inevitable in reinforcement learning (RL) for large language models (LLMs) due to rollout staleness from asynchronous t

VideoRouter: Query-Adaptive Dual Routing for Efficient Long-Video Understanding

SafetyDGX agent

arXiv:2605.05848v2 Announce Type: replace-cross Abstract: Video large multimodal models increasingly face a scalability bottleneck: long videos produce excessively long visual-token sequences, which s

VISD: Enhancing Video Reasoning via Structured Self-Distillation

SafetyDGX agent

arXiv:2605.06094v2 Announce Type: replace-cross Abstract: Training VideoLLMs for complex reasoning remains challenging due to sparse sequence level rewards and the lack of fine grained credit assignme

what i have been saying for 6 years. maybe now you will believe it?

SafetyDGX agent

what i have been saying for 6 years. maybe now you will believe it? 📁 Fei-Fei Li, former Google Chief Scientist, says the industry is dangerously fixated on language models. Most of the real economy i

When Descent Is Too Stable: Event-Triggered Hamiltonian Learning to Optimize

SafetyDGX agent

arXiv:2605.06868v1 Announce Type: new Abstract: Fixed-budget nonconvex optimization can fail not because local descent is unstable, but because it is too stable: after reaching a nearby stationary poi

Which has better odds? Generative AI earning $1.6 trillion/year or the roulette wheel landing on zero?

SafetyDGX agent

Which has better odds? Generative AI earning $1.6 trillion/year or the roulette wheel landing on zero? One estimate of how much annual revenue AI needs to “make sense”: 1.6 trillion. That’s four times

Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

SafetyDGX agent

arXiv:2605.05558v2 Announce Type: replace Abstract: A natural intuition about the economics of AI agents is that, because agents can be replicated at very low marginal cost, agent labor may be supplie

Zero-Shot Imagined Speech Decoding via Imagined-to-Listened MEG Mapping

SafetyDGX agent

arXiv:2605.08075v1 Announce Type: new Abstract: Decoding imagined speech from non-invasive brain recordings is challenging because imagined datasets are scarce and difficult to align temporally across

10 May 2026

32,000 views for my tweet; 2.2 million for the nonsense it critiques. so typical. lies and distortions beating debunking by a factor of roug…

SafetyDGX agent

32,000 views for my tweet; 2.2 million for the nonsense it critiques. so typical. lies and distortions beating debunking by a factor of roughly 100. Wanna get a million views? Make stuff up. Take a ti

a tweet to keep in mind when Sam testifies this coming week

SafetyDGX agent

a tweet to keep in mind when Sam testifies this coming week How can anybody take seriously your claim that “Working towards prosperity for everyone, empowering all people, and advancing science and te

FAR more scary than the misunderstood METR time horizon graph everyone here seems to be (wrongly) panicking about.

SafetyDGX agent

FAR more scary than the misunderstood METR time horizon graph everyone here seems to be (wrongly) panicking about. The #AI circular funding bubble is already twice the size of the outstanding debt in

Happy to put money against superintelligence in 2029.

SafetyDGX agent

Happy to put money against superintelligence in 2029. Why Superintelligence is Alien Intelligence (A Brief Outline of the Future) We are no longer in the realm of normal technological progress. If the

hey @elonmusk, if you really care about making X a source of truth as you used to claim, take note:

SafetyDGX agent

hey @elonmusk, if you really care about making X a source of truth as you used to claim, take note: @GaryMarcus lies are optimized for engagement, truth is optimized for accuracy the algorithm doesnt

Many are valiantly fighting against the age of bullshit. Bullshit is winning, I'm sorry to report. 😠

SafetyDGX agent

Many are valiantly fighting against the age of bullshit. Bullshit is winning, I'm sorry to report. 😠 32,000 views for my tweet; 2.2 million for the nonsense it critiques. so typical. lies and distorti

Oy. According to a new paper in The Lancet, the rate of made-up citations in biomedical papers has increased by more than 12x since 2023. ht…

SafetyDGX agent

Oy. According to a new paper in The Lancet, the rate of made-up citations in biomedical papers has increased by more than 12x since 2023. https://www.thelancet.com/journals/lancet/article/PIIS0140-673

Remarkable numbers here. Goldman Sachs: 'The consensus of analysts is now for the mega-cap US hyperscalers to spend $755 billion on capex in…

SafetyDGX agent

Remarkable numbers here. Goldman Sachs: 'The consensus of analysts is now for the mega-cap US hyperscalers to spend $755 billion on capex in 2026, representing growth of +83% vs. 2025. This capex is e

> shipped the agent > opened the dashboard > latency: fine > error rate: fine > users: unhappy > checked the responses > technically correct…

SafetyDGX agent

> shipped the agent > opened the dashboard > latency: fine > error rate: fine > users: unhappy > checked the responses > technically correct > wrong tool called 3 steps earlier > no trace to follow >

Terence Tao - 'AI tools are like taking a helicopter to drop you off at the site. You miss all the benefits of the journey itself. You just …

SafetyDGX agent

Terence Tao - 'AI tools are like taking a helicopter to drop you off at the site. You miss all the benefits of the journey itself. You just get right to the destination, which actually was only just a

wondering if @embirico has numbers on what % of codex users use this mode and how much it has gone up over the last month its a decent proxy…

SafetyDGX agent

The post asks about user engagement metrics for a specific mode within Codex, requesting data on what percentage of users utilize it and how adoption has changed over the past month. The author sugges

Wow! Wonder how many people on this site hyped this paper – and how many of them will walk back their hype now that the paper has been retra…

SafetyDGX agent

Wow! Wonder how many people on this site hyped this paper – and how many of them will walk back their hype now that the paper has been retracted. A year-old nature paper that advocated the use of Chat

9 May 2026

1. take mostly horizontal part 2. infer vertical

SafetyDGX agent

This post likely discusses a cognitive or computational principle where one should primarily focus on horizontal (broad, general, or foundational) aspects of a problem or system, then use inference or

← Previous
1…179180181182183…240
Next →