AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlog
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,617 results
Industry

Texas AG Ken Paxton sues Netflix for allegedly spying on consumers by collecting their data without consent, and designing its platform to be addictive (Jonathan Stempel/Reuters)

DGX agent

Jonathan Stempel / Reuters: Texas AG Ken Paxton sues Netflix for allegedly spying on consumers by collecting their data without consent, and designing its platform to be addictive — Netflix (NFLX.O) w

industrytechmeme
11 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Safety

TextLDM: Language Modeling with Continuous Latent Diffusion

DGX agent

arXiv:2605.07748v1 Announce Type: new Abstract: Diffusion Transformers (DiT) trained with flow matching in a VAE latent space have unified visual generation across images and videos. A natural next st

safetyarxiv-cs-cl
11 May 2026
Safety

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone f…

DGX agent

That @NIST CAISI announcement from last week about Google DeepMind, xAI and Microsoft signing new deals for pre-deployment testing is gone from their website, link no longer found, and I can't get any

safetygary-marcus--x
11 May 2026
Agents

The AI-Native Large-Scale Agile Software Development Manifesto

DGX agent

arXiv:2605.07717v1 Announce Type: cross Abstract: Despite the widespread adoption of agile methods, achieving true agility at scale remains elusive. Large-scale agile frameworks remain largely human-c

agentsarxiv-cs-ai
11 May 2026
Model Releases

The best way to level up from 1 agent => many agents. No more cycling between terminal tabs

DGX agent

This post discusses strategies for scaling from managing a single AI agent to coordinating multiple agents efficiently, likely addressing workflow challenges and tooling improvements that eliminate th

model-releasesboris-cherny--x
11 May 2026
Agents

The Context Gathering Decision Process: A POMDP Framework for Agentic Search

DGX agent

arXiv:2605.07042v1 Announce Type: new Abstract: Large Language Model (LLM) agents are deployed in complex environments -- such as massive codebases, enterprise databases, and conversational histories

agentsarxiv-cs-ai
11 May 2026
Model Releases

The Convergence Gap: Instruction-Tuned Language Models Stabilize Later in the Forward Pass

DGX agent

arXiv:2605.07282v1 Announce Type: new Abstract: Final outputs hide when a checkpoint commits to its next-token prediction. We introduce the convergence gap, a model-diffing diagnostic that decodes eac

model-releasesarxiv-cs-lg
11 May 2026
Safety

The Cost of Consensus: Malignant Epistemic Herding and Adaptive Gating in Distributed Multi-Agent Search

DGX agent

arXiv:2605.06988v1 Announce Type: cross Abstract: Distributed agents in real-world settings frequently must coordinate under uncertainty with only partial observations. Coordination is necessary to sh

safetyarxiv-cs-ro
11 May 2026
Model Releases

The Coupling Tax: How Shared Token Budgets Undermine Visible Chain-of-Thought Under Fixed Output Limits

DGX agent

arXiv:2605.07686v1 Announce Type: new Abstract: Chain-of-thought reasoning is often treated as a monotone way to improve language-model accuracy by letting a model think longer. We identify a counterv

model-releasesarxiv-cs-lg
11 May 2026
Research

The Download: the hantavirus outbreak and Musk v. Altman week 2

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. Here’s what you need to know about the cruise ship hantavirus

researchmit-tech-review
11 May 2026
Model Releases

The EDelta-MHC-Geo Transformer: Adaptive Geodesic Operations with Guaranteed Orthogonality

DGX agent

arXiv:2605.06729v1 Announce Type: cross Abstract: We present the EDelta-MHC-Geo Transformer, a novel architecture that unifies Manifold-Constrained Hyper-Connections (mHC), Deep Delta Learning (DDL),

model-releasesarxiv-cs-ai
11 May 2026
Safety

The Effect of Mini-Batch Noise on the Implicit Bias of Adam

DGX agent

arXiv:2602.01642v2 Announce Type: replace-cross Abstract: With limited high-quality data and growing compute, multi-epoch training is gaining back its importance across sub-areas of deep learning. Ada

safetyarxiv-cs-ai
11 May 2026
Research

The Effective Depth Paradox: Evaluating the Relationship between Architectural Topology and Trainability in Deep CNNs

DGX agent

arXiv:2602.13298v3 Announce Type: replace-cross Abstract: This paper investigates the relationship between convolutional neural network (CNN) topology and image recognition performance through a compa

researcharxiv-cs-ai
11 May 2026
Safety

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

DGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

safetyarxiv-cs-ai
11 May 2026
Applications

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do…

DGX agent

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do science & the same-y writing limits their usefulness in man

applicationsethan-mollick--x
11 May 2026
Tools

the inside story of the legendary Cog House. i believe there have not been any public photos of this place until now (bc we were explicitly …

DGX agent

the inside story of the legendary Cog House. i believe there have not been any public photos of this place until now (bc we were explicitly not allowed to lol) as an advisor its been awe inspiring to

toolsswyx--x
11 May 2026
Agents

The internet gave language models their data for free. Robots don’t have that. Every trajectory has to be earned through hardware, time, tel…

DGX agent

The internet gave language models their data for free. Robots don’t have that. Every trajectory has to be earned through hardware, time, teleoperators, and real consequences. Shrey’s piece on simulati

agentsyohei-nakajima--x
11 May 2026
Safety

The Limits of AI-Driven Allocation: Optimal Screening under Aleatoric Uncertainty

DGX agent

arXiv:2605.07979v1 Announce Type: new Abstract: The rise of machine learning has shifted targeted resource allocation in policy and humanitarian settings toward algorithmic targeting based on predicte

safetyarxiv-cs-ai
11 May 2026
Model Releases

The Memory Curse: How Expanded Recall Erodes Cooperative Intent in LLM Agents

DGX agent

arXiv:2605.08060v1 Announce Type: cross Abstract: Context window expansion is often treated as a straightforward capability upgrade for LLMs, but we find it systematically fails in multi-agent social

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

// The Memory Curse in LLM Agents // (bookmark it) Long histories apparently degrades agents as they become increasingly history-following a…

DGX agent

// The Memory Curse in LLM Agents // (bookmark it) Long histories apparently degrades agents as they become increasingly history-following and risk-minimizing. Across 7 LLMs and 4 social dilemma games

tutorialsdair-ai--x
11 May 2026
Research

The Minimax Rate of Second-Order Calibration

DGX agent

arXiv:2605.07808v1 Announce Type: new Abstract: We characterize the minimax rate of estimating the second-order calibration error for binary classification, which quantifies whether a higher-order pre

researcharxiv-cs-lg
11 May 2026
Safety

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

DGX agent

arXiv:2605.07462v1 Announce Type: cross Abstract: Moltbook is a Reddit-like platform where OpenClaw agents post, comment, and vote at scale - a so far unprecedented incident that comes with serious sa

safetyarxiv-cs-ai
11 May 2026
Applications

The new AI-powered Google Finance is expanding to Europe.

DGX agent

Google's AI-powered Google Finance is launching across Europe this week with full local language support. The reimagined platform offers capabilities including AI-powered research that lets users ask

applicationsgoogle-ai
11 May 2026
Tools

The next generation of models won't just generate images - they'll understand worlds, motion, interaction, and action. We've been building t…

DGX agent

The next generation of models won't just generate images - they'll understand worlds, motion, interaction, and action. We've been building toward this for a while. Visual intelligence is becoming real

toolsswyx--x
11 May 2026
Model Releases

The Position Curse: LLMs Struggle to Locate the Last Few Items in a List

DGX agent

arXiv:2605.07127v1 Announce Type: cross Abstract: Modern large language models (LLMs) can find a needle in a haystack (locating a single relevant fact buried among hundreds of thousands of irrelevant

model-releasesarxiv-cs-cl
11 May 2026
Safety

The Proxy Presumption: From Semantic Embeddings to Valid Social Measures

DGX agent

arXiv:2605.07409v1 Announce Type: new Abstract: Natural Language Processing is rapidly evolving into a primary instrument for Computational Social Science, with researchers increasingly using embeddin

safetyarxiv-cs-cl
11 May 2026
Model Releases

The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking

DGX agent

arXiv:2605.06707v1 Announce Type: cross Abstract: This paper presents an eight-week observational comparison of 68 single-file HTML generations collected across 17 public experiments in the 'HTML AI B

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Text Uncanny Valley: Non-Monotonic Performance Degradation in LLM Information Retrieval

DGX agent

arXiv:2605.07186v1 Announce Type: cross Abstract: Existing Large Language Model (LLM) benchmarks primarily focus on syntactically correct inputs, leaving a significant gap in evaluation on imperfect t

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Translation Tax Is Not a Scalar: A Counterfactual Audit of English-Source Cue Inheritance in Chinese Multilingual Benchmarks

DGX agent

arXiv:2605.07093v1 Announce Type: cross Abstract: The Translation Tax is often treated as a scalar: translated benchmarks are assumed to inflate scores by preserving English-source cues. We audit this

model-releasesarxiv-cs-ai
11 May 2026
Applications

The US Commerce Department removed from its website details about its May 5 agreement with Google, xAI, and Microsoft to test their AI models (Courtney Rozen/Reuters)

DGX agent

Courtney Rozen / Reuters: The US Commerce Department removed from its website details about its May 5 agreement with Google, xAI, and Microsoft to test their AI models — The U.S. Commerce Department r

applicationstechmeme
11 May 2026
Safety

Theoretical Limits of Language Model Alignment

DGX agent

arXiv:2605.07105v1 Announce Type: cross Abstract: Language model (LM) alignment improves model outputs to reflect human preferences while preserving the capabilities of the base model. The most common

safetyarxiv-cs-cl
11 May 2026
Model Releases

Theory of Optimal Learning Rate Schedules and Scaling Laws for a Random Feature Model

DGX agent

arXiv:2602.04774v2 Announce Type: replace-cross Abstract: Setting the learning rate (LR) for a deep learning model is a critical part of successful training. Choosing LRs is often done empirically wit

model-releasesarxiv-cs-lg
11 May 2026
Local Ai

Think Before You Drive: World Model-Inspired Multimodal Grounding for Autonomous Vehicles

DGX agent

arXiv:2512.03454v4 Announce Type: replace-cross Abstract: Interpreting natural-language commands to localize target objects is critical for autonomous driving (AD). Existing visual grounding (VG) meth

local-aiarxiv-cs-ai
11 May 2026
Safety

Think-with-Rubrics: From External Evaluator to Internal Reasoning Guidance

DGX agent

arXiv:2605.07461v1 Announce Type: new Abstract: Rubrics have been extensively utilized for evaluating unverifiable, open-ended tasks, with recent research incorporating them into reward systems for re

safetyarxiv-cs-cl
11 May 2026
Industry

Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration (Thinking Machines Lab)

DGX agent

Thinking Machines Lab: Thinking Machines Lab details interaction models, which can think and respond in real time, letting users and AI interact continuously for better collaboration — Today, we're an

industrytechmeme
11 May 2026
Model Releases

THINKSAFE: Self-Generated Safety Alignment for Reasoning Models

DGX agent

arXiv:2601.23143v2 Announce Type: replace Abstract: Large reasoning models (LRMs) achieve remarkable performance by leveraging reinforcement learning (RL) on reasoning tasks to generate long chain-of-

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ThinKV: Thought-Adaptive KV Cache Compression for Efficient Reasoning Models

DGX agent

arXiv:2510.01290v2 Announce Type: replace Abstract: The long-output context generation of large reasoning models enables extended chain of thought (CoT) but also drives rapid growth of the key-value (

model-releasesarxiv-cs-lg
11 May 2026
Research

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters …

DGX agent

Thinky's secret plan: 1: Increase Human<->AI bandwidth 2: Raise ceiling of human+AI intelligence 3: Help humans continue as main-characters in the new world We are at Step 1. Interaction Models are gr

researchsoumith-chintala--x
11 May 2026
Agents

This AI watches its own codebase, flags missing monitors, and opens PRs to fix bugs it finds. @Shevchenkoaalex on @TryRamp’s self-monitoring…

DGX agent

This post describes an AI system that monitors its own codebase to identify missing monitoring points and automatically generates pull requests to fix detected bugs, demonstrating autonomous self-impr

agentsharrison-chase--x
11 May 2026
Tutorials

This is a tutorial on diffusion and flow matching, based on my previous postings here. I’ve made available the PDF, a python notebook for pe…

DGX agent

This is a tutorial on diffusion and flow matching, based on my previous postings here. I’ve made available the PDF, a python notebook for people to play with it, and the TEX source so hopefully one of

tutorialsemad-mostaque--x
11 May 2026
Applications

This is going to get even worse as people realize that careful tuning in their prompts can make AI writing seem not like AI writing to reade…

DGX agent

This is going to get even worse as people realize that careful tuning in their prompts can make AI writing seem not like AI writing to readers. We expect word counts to align, in some way, with thinki

applicationsethan-mollick--x
11 May 2026
Agents

This is harder to build than it looks. Preserving full conversational context while swapping underlying model providers mid-flight is a surp…

DGX agent

This is harder to build than it looks. Preserving full conversational context while swapping underlying model providers mid-flight is a surprisingly deep systems problem. Most tools drop state or forc

agentsharrison-chase--x
11 May 2026
Safety

This is what a useless hype lifecycle looks like.

DGX agent

Gary Marcus critiques the typical hype cycle pattern where emerging technologies experience inflated expectations followed by inevitable disappointment. The post likely illustrates this cycle using a

safetygary-marcus--x
11 May 2026
Agents

This seems like a critical reason to open up about AI use in academia. Scholars are using old AI models, badly, and not talking about it. Ne…

DGX agent

This seems like a critical reason to open up about AI use in academia. Scholars are using old AI models, badly, and not talking about it. New models hallucinate very few citations, and good agentic ha

agentsethan-mollick--x
11 May 2026
Research

This works really well btw, at the end of your query ask your LLM to 'structure your response as HTML', then view the generated file in your…

DGX agent

This works really well btw, at the end of your query ask your LLM to 'structure your response as HTML', then view the generated file in your browser. I've also had some success asking the LLM to prese

researchkarpathy--x
11 May 2026
Agents

Thoughts on GitLab's workforce reduction' and 'structural and strategic decisions'

DGX agent

GitLab Act 2 There's a lot going on in this announcement from GitLab about the 'workforce reduction' and 'structural and strategic decisions' they are making with respect to the agentic era. They're '

agentssimon-willison
11 May 2026
Research

Three-in-One World Model: Energy-Based Consistency, Prediction, and Counterfactual Inference for Marketing Intervention

DGX agent

arXiv:2605.07199v1 Announce Type: new Abstract: Marketing decisions reflect the interaction of latent consumer heterogeneity, time-varying internal states, and explicit interventions, a structure that

researcharxiv-cs-ai
11 May 2026
← Previous
1…13201321132213231324…1826
Next →