AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,813 results
Safety

Temporal Smoothness Doubly Robust Learning for Debiased Knowledge Tracing

DGX agent

arXiv:2605.05958v2 Announce Type: replace Abstract: Knowledge Tracing (KT) is fundamental to intelligent education systems, yet relies on educational logs that are selectively observed. The non-random

safetyarxiv-cs-ai
11 May 2026
Research

Tessellations of Semi-Discrete Flow Matching

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.07513v1 Announce Type: new Abstract: We study Flow Matching in a semi-discrete setting where a Gaussian source is transported toward a discrete target supported on finitely many points. Thi

researcharxiv-cs-lg
11 May 2026
Local Ai

Test-Time Compositional Generalization in Diffusion Models via Concept Discovery

DGX agent

arXiv:2605.07078v1 Announce Type: new Abstract: Compositional generalization requires models to produce novel configurations from familiar parts. In diffusion models, prior compositional generation me

local-aiarxiv-cs-lg
11 May 2026
Model Releases

Test-Time Compute Games

DGX agent

arXiv:2601.21839v2 Announce Type: replace-cross Abstract: Test-time compute has emerged as a promising strategy to enhance the reasoning abilities of large language models (LLMs). However, this strate

model-releasesarxiv-cs-ai
11 May 2026
Research

Testing Noise Assumptions of Learning Algorithms

DGX agent

arXiv:2501.09189v3 Announce Type: replace Abstract: We pose a fundamental question in computational learning theory: can we efficiently test whether a training set satisfies the assumptions of a given

researcharxiv-cs-lg
11 May 2026
Industry

Texas AG Ken Paxton sues Netflix for allegedly spying on consumers by collecting their data without consent, and designing its platform to be addictive (Jonathan Stempel/Reuters)

DGX agent

Jonathan Stempel / Reuters: Texas AG Ken Paxton sues Netflix for allegedly spying on consumers by collecting their data without consent, and designing its platform to be addictive — Netflix (NFLX.O) w

industrytechmeme
11 May 2026
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Safety

TextLDM: Language Modeling with Continuous Latent Diffusion

DGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

safetyarxiv-cs-cl
11 May 2026
Safety

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone f…

DGX agent

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone from their website, link no longer found, and I can't get any

safetygary-marcus--x
11 May 2026
Agents

The AI-Native Large-Scale Agile Software Development Manifesto

DGX agent

arXiv:2605.07717v1 Announce Type: cross Abstract: Despite the widespread adoption of agile methods, achieving true agility at scale remains elusive. Large-scale agile frameworks remain largely human-c

agentsarxiv-cs-ai
11 May 2026
Model Releases

The best way to level up from 1 agent => many agents. No more cycling between terminal tabs

DGX agent

This post discusses strategies for scaling from managing a single AI agent to coordinating multiple agents efficiently, likely addressing workflow challenges and tooling improvements that eliminate th

model-releasesboris-cherny--x
11 May 2026
Agents

The Context Gathering Decision Process: A POMDP Framework for Agentic Search

DGX agent

arXiv:2605.07042v1 Announce Type: new Abstract: Large Language Model (LLM) agents are deployed in complex environments -- such as massive codebases, enterprise databases, and conversational histories

agentsarxiv-cs-ai
11 May 2026
Model Releases

The Convergence Gap: Instruction-Tuned Language Models Stabilize Later in the Forward Pass

DGX agent

arXiv:2605.07282v1 Announce Type: new Abstract: Final outputs hide when a checkpoint commits to its next-token prediction. We introduce the convergence gap, a model-diffing diagnostic that decodes eac

model-releasesarxiv-cs-lg
11 May 2026
Safety

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

DGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

safetyarxiv-cs-ro
11 May 2026
Model Releases

The Coupling Tax: How Shared Token Budgets Undermine Visible Chain-of-Thought Under Fixed Output Limits

DGX agent

arXiv:2605.07686v1 Announce Type: new Abstract: Chain-of-thought reasoning is often treated as a monotone way to improve language-model accuracy by letting a model think longer. We identify a counterv

model-releasesarxiv-cs-lg
11 May 2026
Research

The Download: the hantavirus outbreak and Musk v. Altman week 2

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Here’s what you need to know about the cruise ship hantavirus

researchmit-tech-review
11 May 2026
Model Releases

The EDelta-MHC-Geo Transformer: Adaptive Geodesic Operations with Guaranteed Orthogonality

DGX agent

arXiv:2605.06729v1 Announce Type: cross Abstract: We present the EDelta-MHC-Geo Transformer, a novel architecture that unifies Manifold-Constrained Hyper-Connections (mHC), Deep Delta Learning (DDL),

model-releasesarxiv-cs-ai
11 May 2026
Safety

The Effect of Mini-Batch Noise on the Implicit Bias of Adam

DGX agent

arXiv:2602.01642v2 Announce Type: replace-cross Abstract: With limited high-quality data and growing compute, multi-epoch training is gaining back its importance across sub-areas of deep learning. Ada

safetyarxiv-cs-ai
11 May 2026
Research

The Effective Depth Paradox: Evaluating the Relationship between Architectural Topology and Trainability in Deep CNNs

DGX agent

arXiv:2602.13298v3 Announce Type: replace-cross Abstract: This paper investigates the relationship between convolutional neural network (CNN) topology and image recognition performance through a compa

researcharxiv-cs-ai
11 May 2026
Safety

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

DGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

safetyarxiv-cs-ai
11 May 2026
Applications

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do…

DGX agent

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do science & the same-y writing limits their usefulness in man

applicationsethan-mollick--x
11 May 2026
Tools

the inside story of the legendary Cog House. i believe there have not been any public photos of this place until now (bc we were explicitly …

DGX agent

the inside story of the legendary Cog House. i believe there have not been any public photos of this place until now (bc we were explicitly not allowed to lol) as an advisor its been awe inspiring to

toolsswyx--x
11 May 2026
Agents

The internet gave language models their data for free. Robots don’t have that. Every trajectory has to be earned through hardware, time, tel…

DGX agent

The internet gave language models their data for free. Robots don’t have that. Every trajectory has to be earned through hardware, time, teleoperators, and real consequences. Shrey’s piece on simulati

agentsyohei-nakajima--x
11 May 2026
Safety

The Limits of AI-Driven Allocation: Optimal Screening under Aleatoric Uncertainty

DGX agent

arXiv:2605.07979v1 Announce Type: new Abstract: The rise of machine learning has shifted targeted resource allocation in policy and humanitarian settings toward algorithmic targeting based on predicte

safetyarxiv-cs-ai
11 May 2026
Model Releases

The Memory Curse: How Expanded Recall Erodes Cooperative Intent in LLM Agents

DGX agent

arXiv:2605.08060v1 Announce Type: cross Abstract: Context window expansion is often treated as a straightforward capability upgrade for LLMs, but we find it systematically fails in multi-agent social

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

// The Memory Curse in LLM Agents // (bookmark it) Long histories apparently degrades agents as they become increasingly history-following a…

DGX agent

// The Memory Curse in LLM Agents // (bookmark it) Long histories apparently degrades agents as they become increasingly history-following and risk-minimizing. Across 7 LLMs and 4 social dilemma games

tutorialsdair-ai--x
11 May 2026
Research

The Minimax Rate of Second-Order Calibration

DGX agent

arXiv:2605.07808v1 Announce Type: new Abstract: We characterize the minimax rate of estimating the second-order calibration error for binary classification, which quantifies whether a higher-order pre

researcharxiv-cs-lg
11 May 2026
Safety

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

DGX agent

arXiv:2605.07462v1 Announce Type: cross Abstract: Moltbook is a Reddit-like platform where OpenClaw agents post, comment, and vote at scale - a so far unprecedented incident that comes with serious sa

safetyarxiv-cs-ai
11 May 2026
Applications

The new AI-powered Google Finance is expanding to Europe.

DGX agent

Google's AI-powered Google Finance is launching across Europe this week with full local language support. The reimagined platform offers capabilities including AI-powered research that lets users ask

applicationsgoogle-ai
11 May 2026
Tools

The next generation of models won't just generate images - they'll understand worlds, motion, interaction, and action. We've been building t…

DGX agent

The next generation of models won't just generate images - they'll understand worlds, motion, interaction, and action. We've been building toward this for a while. Visual intelligence is becoming real

toolsswyx--x
11 May 2026
Model Releases

The Position Curse: LLMs Struggle to Locate the Last Few Items in a List

DGX agent

arXiv:2605.07127v1 Announce Type: cross Abstract: Modern large language models (LLMs) can find a needle in a haystack (locating a single relevant fact buried among hundreds of thousands of irrelevant

model-releasesarxiv-cs-cl
11 May 2026
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Model Releases

The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking

DGX agent

arXiv:2605.06707v1 Announce Type: cross Abstract: This paper presents an eight-week observational comparison of 68 single-file HTML generations collected across 17 public experiments in the 'HTML AI B

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Text Uncanny Valley: Non-Monotonic Performance Degradation in LLM Information Retrieval

DGX agent

arXiv:2605.07186v1 Announce Type: cross Abstract: Existing Large Language Model (LLM) benchmarks primarily focus on syntactically correct inputs, leaving a significant gap in evaluation on imperfect t

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Translation Tax Is Not a Scalar: A Counterfactual Audit of English-Source Cue Inheritance in Chinese Multilingual Benchmarks

DGX agent

arXiv:2605.07093v1 Announce Type: cross Abstract: The Translation Tax is often treated as a scalar: translated benchmarks are assumed to inflate scores by preserving English-source cues. We audit this

model-releasesarxiv-cs-ai
11 May 2026
Applications

The US Commerce Department removed from its website details about its May 5 agreement with Google, xAI, and Microsoft to test their AI models (Courtney Rozen/Reuters)

DGX agent

Courtney Rozen / Reuters: The US Commerce Department removed from its website details about its May 5 agreement with Google, xAI, and Microsoft to test their AI models — The U.S. Commerce Department r

applicationstechmeme
11 May 2026
Safety

Theoretical Limits of Language Model Alignment

DGX agent

arXiv:2605.07105v1 Announce Type: cross Abstract: Language model (LM) alignment improves model outputs to reflect human preferences while preserving the capabilities of the base model. The most common

safetyarxiv-cs-cl
11 May 2026
Model Releases

Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model

DGX agent

arXiv:2602.04774v2 Announce Type: replace-cross Abstract: Setting the learning rate (LR) for a deep learning model is a critical part of successful training. Choosing LRs is often done empirically wit

model-releasesarxiv-cs-lg
11 May 2026
Local Ai

Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles

DGX agent

arXiv:2512.03454v4 Announce Type: replace-cross Abstract: Interpreting natural-language commands to localize target objects is critical for autonomous driving (AD). Existing visual grounding (VG) meth

local-aiarxiv-cs-ai
11 May 2026
Safety

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

DGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

safetyarxiv-cs-cl
11 May 2026
Industry

Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration (Thinking Machines Lab)

DGX agent

Thinking Machines Lab: Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration — Today, we're an

industrytechmeme
11 May 2026
Model Releases

THINKSAFE: Self-Generated Safety Alignment for Reasoning Models

DGX agent

arXiv:2601.23143v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve remarkable performance by leveraging reinforcement learning (RL) on reasoning tasks to generate long chain-of-

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

DGX agent

arXiv:2510.01290v2 Announce Type: replace Abstract: The long-output context generation of large reasoning models enables extended chain of thought (CoT) but also drives rapid growth of the key-value (

model-releasesarxiv-cs-lg
11 May 2026
Research

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters …

DGX agent

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters in the new world We are at Step 1. Interaction Models are gr

researchsoumith-chintala--x
11 May 2026
Agents

This AI watches its own codebase, flags missing monitors, and opens PRs to fix bugs it finds. @Shevchenkoaalex on @TryRamp’s self-monitoring…

DGX agent

This post describes an AI system that monitors its own codebase to identify missing monitoring points and automatically generates pull requests to fix detected bugs, demonstrating autonomous self-impr

agentsharrison-chase--x
11 May 2026
Tutorials

This is a tutorial on diffusion and flow matching, based on my previous postings here. I’ve made available the PDF, a python notebook for pe…

DGX agent

This is a tutorial on diffusion and flow matching, based on my previous postings here. I’ve made available the PDF, a python notebook for people to play with it, and the TEX source so hopefully one of

tutorialsemad-mostaque--x
11 May 2026
Applications

This is going to get even worse as people realize that careful tuning in their prompts can make AI writing seem not like AI writing to reade…

DGX agent

This is going to get even worse as people realize that careful tuning in their prompts can make AI writing seem not like AI writing to readers. We expect word counts to align, in some way, with thinki

applicationsethan-mollick--x
11 May 2026
Agents

This is harder to build than it looks. Preserving full conversational context while swapping underlying model providers mid-flight is a surp…

DGX agent

This is harder to build than it looks. Preserving full conversational context while swapping underlying model providers mid-flight is a surprisingly deep systems problem. Most tools drop state or forc

agentsharrison-chase--x
11 May 2026
← Previous
1…13241325132613271328…1830
Next →