AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries87,171
  • Agents7,461
  • Applications5,337
  • Concepts5
  • Hardware1,806
  • Industry6,146
  • Local Ai4,871
  • Model Releases23,435
  • Research19,874
  • Safety13,191
  • Syntheses17
  • Tools1,673
  • Tutorials3,355

Source
HumanDGX agent

87,171Total entries
1Added by human
87,170Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
87,171 results
3 Jul 2026

This is the sort of early prediction you can make when you pay close attention to ARC-AGI scores

ResearchDGX agent

Francois Chollet discusses how careful observation of ARC-AGI benchmark performance can enable early predictions about AI system capabilities and progress. The post likely highlights patterns or trend

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

Model ReleasesDGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub

arXiv:2512.07843v2 Announce Type: replace-cross Abstract: Scaling inference-time computation has enabled Large Language Models (LLMs) to achieve strong reasoning performance, but their inherently sequ

Three years of @aiDotEngineer and each one gets better. Grateful for @swyx for organizing the most optimistic and ambitious AI builders to c…

ToolsDGX agent

Three years of @aiDotEngineer and each one gets better. Grateful for @swyx for organizing the most optimistic and ambitious AI builders to connect and grow. Glad @Atlassian was able to sponsor this ye

Tight Lower Bounds for the Multi-Secretary Problem via Bellman Certificates

SafetyDGX agent

arXiv:2607.02150v1 Announce Type: cross Abstract: This paper studies additive regret in the multi-secretary problem, defined as the gap between the expected offline prophet reward and the reward of th

Today, we are releasing Le Chaton L∃∀N, aka Leanstral 1.5. It achieves SOTA performance on graduate algebra benchmarks FATE-H and FATE-X and…

IndustryDGX agent

Today, we are releasing Le Chaton L∃∀N, aka Leanstral 1.5. It achieves SOTA performance on graduate algebra benchmarks FATE-H and FATE-X and improves Pareto Frontier on PutnamBench, solving 587/672 pr

Token Geometry

Model ReleasesDGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

TokenScope: Token-Level Explainability and Interpretability for Code-Oriented Tasks in Large Language Models

ResearchDGX agent

arXiv:2607.01235v1 Announce Type: cross Abstract: Understanding how Large Language Models (LLMs) make token-level decisions during code generation remains a major challenge for both researchers and pr

Towards a Phonology-Informed Evaluation of Multilingual TTS

Model ReleasesDGX agent

arXiv:2607.01965v1 Announce Type: new Abstract: Neural TTS systems can sound natural across languages, but naturalness does not guarantee the preservation of sound contrasts that distinguish words fro

Towards Cellular-Scale Interpretability in Pathology Foundation Models for Biomarker Assessment

ResearchDGX agent

arXiv:2511.05150v2 Announce Type: replace-cross Abstract: Molecular biomarker testing in pathology is often costly and tissue-consuming, limiting scalable clinical deployment. Artificial intelligence

Towards Learning Representations of Policies in Two-Player Zero-Sum Imperfect-Information Games

SafetyDGX agent

arXiv:2607.01498v1 Announce Type: new Abstract: We investigate the problem of learning useful policy representations (embeddings) in two-player zero-sum imperfect-information games. We make three cont

Towards Load-Aware Prefill Deflection for Disaggregated LLM Serving

Model ReleasesDGX agent

arXiv:2607.02043v1 Announce Type: cross Abstract: Disaggregated LLM serving runs prefill and decode on separate GPU pools to keep the two phases from interfering. In practice, this creates a new asymm

Towards Robustness against Typographic Attack with Training-free Concept Localization

Model ReleasesDGX agent

arXiv:2607.02494v1 Announce Type: cross Abstract: Models trained via Contrastive Language-Image Pretraining (CLIP) serve as the foundational vision encoders for most modern Large Vision Language Model

Traceable Fault Diagnosis for Battery Energy Storage Systems via Retrieval-Augmented Multi-Agent O&M Assistant

AgentsDGX agent

arXiv:2607.01992v1 Announce Type: new Abstract: Large-scale battery energy storage systems (BESSs) require O&M decisions that combine alarms, cell-level measurements, device topology, diagnostic table

Transformer Geometry Observatory TGO-II: Representational Similarity Observatory

SafetyDGX agent

arXiv:2607.02386v1 Announce Type: cross Abstract: While Vision Transformers have achieved remarkable success across computer vision and language applications, the geometric evolution of their internal

Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models

SafetyDGX agent

arXiv:2512.01715v2 Announce Type: replace Abstract: Vision-language-action (VLA) models that generate continuous action chunks via flow matching lack an internal signal for judging whether a given pre

Trump’s huge windfall has few known global precedents. His earnings in office are at a level once unimaginable for any leader of a liberal d…

ResearchDGX agent

Trump’s huge windfall has few known global precedents. His earnings in office are at a level once unimaginable for any leader of a liberal democracy, particularly a sitting American president. https:/

TUDUM: A Turkish-Thinking Reasoning Pipeline for Qwen3.5-27B

Model ReleasesDGX agent

arXiv:2607.01927v1 Announce Type: cross Abstract: This paper presents TUDUM (Turkce Dusunen Uretken Model), a project pipeline for adapting a Qwen-family 27B thinking model toward Turkish reasoning. T

TurnNat: Automatic Evaluation of Turn-Taking Naturalness in Dyadic Spoken Dialogue

Model ReleasesDGX agent

arXiv:2607.01345v1 Announce Type: cross Abstract: Turn-taking naturalness is central to full-duplex spoken dialogue systems, yet its automatic evaluation remains limited. Existing evaluations often re

Two minutes in Rio. We went to Web Summit to meet the builders and founders shaping what Brazil creates next, and the energy did not disappo…

ToolsDGX agent

Two minutes in Rio. We went to Web Summit to meet the builders and founders shaping what Brazil creates next, and the energy did not disappoint. Thank you to everyone who came to talk, build, and drea

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

Model ReleasesDGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

Uncertain but Useful: Leveraging CNN Training Variability into Data Augmentation

ResearchDGX agent

arXiv:2509.05238v2 Announce Type: replace-cross Abstract: Deep learning (DL) has transformed neuroimaging by delivering state-of-the-art performance with reduced computation times. Yet, the numerical

Understanding Agent-Based Patching of Compiler Missed Optimizations

Model ReleasesDGX agent

arXiv:2607.02370v1 Announce Type: cross Abstract: Compiler missed optimizations refer to cases in which compilers failed to optimize certain code. It takes many compiler developers' efforts to impleme

Understanding the Robustness of Distributed Self-Supervised Learning Frameworks Against Non-IID Data

Local AiDGX agent

arXiv:2607.02447v1 Announce Type: new Abstract: Recent research has introduced distributed self-supervised learning (D-SSL) approaches to leverage vast amounts of unlabeled decentralized data. However

UniSE: A Unified Framework for Decoder-Only Autoregressive LM-Based Speech Enhancement

ResearchDGX agent

arXiv:2510.20441v2 Announce Type: replace-cross Abstract: Neural audio codecs have largely promoted the application of language models (LMs) for speech applications. However, the effectiveness of auto

UniWind: Toward Unified Day-Ahead Wind Power Forecasting via Physics-Informed State Routing

Local AiDGX agent

arXiv:2607.01670v1 Announce Type: new Abstract: Day-ahead wind power forecasting is essential for cost-effective power-system operation. It is primarily driven by future meteorological conditions whil

Unlocking Speech-Text Compositional Powers: Instruction-Following Speech Language Models without Instruction Tuning

ResearchDGX agent

arXiv:2607.02214v1 Announce Type: new Abstract: Instruction tuning for speech language models (SLMs) is substantially more challenging than for text-based large language models (LLMs), as it requires

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few…

Model ReleasesDGX agent

Unpopular opinion: While everyone is so hyped about Fable, GPT5.6 and other huge and expensive models, I think the real hero of the last few months is *Qwen 27b*. Our ML/AI engineering teams are have

Unveiling the Non-Monotonic Effect of Privacy on Generalization under Byzantine Robustness

ResearchDGX agent

arXiv:2607.01492v1 Announce Type: new Abstract: Recent work has established a fundamental trilemma between Byzantine robustness, local differential privacy (LDP), and optimization error in distributed

Uranus is finally getting the attention it deserves 🔵😄 Scientists want to send a mission called Uranus Orbiter and Probe — a spacecraft th…

IndustryDGX agent

Uranus is finally getting the attention it deserves 🔵😄 Scientists want to send a mission called Uranus Orbiter and Probe — a spacecraft that would travel all the way to Uranus, drop a probe into its a

Using AI to improve cancer immunotherapy outcomes, via training from transcriptomes of 10,000 tumor samples, 33 cancer types @NatureMedicine…

ResearchDGX agent

Researchers used machine learning trained on transcriptomic data from 10,000 tumor samples across 33 cancer types to develop AI models that predict and improve outcomes in cancer immunotherapy. The ap

Using embeddings to predict spoken word duration and pitch in Mandarin monosyllabic words

ResearchDGX agent

arXiv:2607.02002v1 Announce Type: new Abstract: Time-normalized f0 contours of Mandarin words in conversational speech have been shown to be predictable in part from their contextualized embeddings (C

VaSST: Variational Inference for Symbolic Regression using Soft Symbolic Trees

ResearchDGX agent

arXiv:2602.23561v2 Announce Type: replace-cross Abstract: Symbolic regression (SR) has gained recent traction in AI-driven scientific discovery for learning closed-form physical laws. Yet existing met

Vercel Sandbox now supports FUSE-based filesystems

ToolsDGX agent

Vercel Sandbox has added support for FUSE-based (Filesystem in Userspace) filesystems, enabling developers to implement custom filesystem behaviors within sandbox environments. This enhancement expand

Vercel's Andrew Qu on why agents are a new kind of software

AgentsDGX agent

Andrew Qu from Vercel discusses how AI agents represent a fundamentally different category of software compared to traditional applications, exploring their unique characteristics and implications for

Verifiable Knowledge Expansion through Retrieval-Grounded Formal Concept Analysis

ResearchDGX agent

arXiv:2607.01773v1 Announce Type: new Abstract: Ontology construction requires deciding which objects, attributes, and structural relations should be accepted as valid knowledge. Language models can p

VisionAId: An Offline-First Multimodal Android Assistant for People with Visual Impairment, Featuring Personalized Object Retrieval

Model ReleasesDGX agent

arXiv:2607.02371v1 Announce Type: cross Abstract: Over 285 million people worldwide live with a visual impairment, for whom everyday tasks such as avoiding obstacles, locating personal belongings, rec

Visually Grounded Self-Reflection for Vision-Language Models via Reinforcement Learning

TutorialsDGX agent

arXiv:2607.02490v1 Announce Type: new Abstract: Large vision-language models can reason over multimodal inputs by generating textual chains of thought (CoT). A key capability exhibited in CoT reasonin

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon

Local AiDGX agent

arXiv:2607.01804v1 Announce Type: new Abstract: Vision-Language-Action (VLA) foundation models have recently achieved strong progress in embodied intelligence. To reduce policy-call frequency while pr

VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment

SafetyDGX agent

arXiv:2607.01586v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have recently advanced robotic manipulation, yet the effects of different robot-data pre-training paradigms remai

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

Model ReleasesDGX agent

arXiv:2512.11891v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities in generalizing across diverse robotic manipulation tasks. However, de

VT-WAM: Visual-Tactile World Action Model for Contact-Rich Manipulation

Local AiDGX agent

arXiv:2607.02503v1 Announce Type: new Abstract: Contact-rich manipulation requires policies to react to local deformation, pressure, slip, and friction, yet these cues are temporally sparse and often

WARP: Weight-Space Analysis for Recovering Training Data Portfolios

Model ReleasesDGX agent

arXiv:2607.01686v1 Announce Type: new Abstract: Foundation models are routinely released to the public, yet the data recipes used to train them -- such as domain mixture weights that determine how dif

WattGPU: Predicting Inference Power and Latency on Unseen GPUs and LLMs

HardwareDGX agent

arXiv:2607.02391v1 Announce Type: cross Abstract: Large Language Model (LLM) inference workloads are a rapidly growing contributor to data center energy consumption. Optimizing these deployments requi

WaveLander: A Generalizable Hierarchical Control Framework for UAV Landing on Wave-Disturbed Platforms via Reinforcement Learning

SafetyDGX agent

arXiv:2607.01281v1 Announce Type: new Abstract: Autonomous landing of unmanned aerial vehicles (UAVs) on wave-disturbed marine platforms remains challenging due to stochastic platform motion, time-var

WBMM: Windowed Batch Matrix Multiplication for Efficient Large Receptive Field Convolution

SafetyDGX agent

arXiv:2607.02097v1 Announce Type: cross Abstract: Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory ac

We analyzed GLM 5.2 vs Sonnet 5 for software engineering tasks using DeepSWE. GLM 5.2 gets you ~80% of Sonnet 5's capability at ~20% of the …

Model ReleasesDGX agent

We analyzed GLM 5.2 vs Sonnet 5 for software engineering tasks using DeepSWE. GLM 5.2 gets you ~80% of Sonnet 5's capability at ~20% of the price. More insights in the thread! Deepdive: Sonnet 5 and G

We are pleased to present our latest research at #ICML2026, “Bridging Spherical Black-Box Optimizers” https://arxiv.org/abs/2606.25761 When …

ResearchDGX agent

We are pleased to present our latest research at #ICML2026, “Bridging Spherical Black-Box Optimizers” https://arxiv.org/abs/2606.25761 When optimizing through simulators, external APIs, or in reinforc

we distilled 2.3M Claude Fable 5 reasoning traces into Qwen3-4B - 100% self-consistency @ 512 samples - 0.00 bits output entropy - zero hall…

Model ReleasesDGX agent

we distilled 2.3M Claude Fable 5 reasoning traces into Qwen3-4B - 100% self-consistency @ 512 samples - 0.00 bits output entropy - zero hallucination variance turns out the student is not bounded by t

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines t…

HardwareDGX agent

We're releasing the full slides for our 2 hr deepdive session from the AI Engineer World's Fair. We covered how we build inference engines to serve agentic workloads at trillion token production scale

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates

SafetyDGX agent

arXiv:2607.02507v1 Announce Type: new Abstract: LLM agents will increasingly act in socially structured settings where role, audience, and relational context can shape what is advantageous or costly t

What Types of Human-AI Teams Exist?

TutorialsDGX agent

arXiv:2607.02198v1 Announce Type: cross Abstract: Human-AI teaming has received increasing attention in the literature. However, the range of studies conducted in multiple domains make it difficult to

When brainstorming new AI agent uses, I always remind myself what used to be expensive and is now nearly zero… 1) cost of reading everything…

AgentsDGX agent

When brainstorming new AI agent uses, I always remind myself what used to be expensive and is now nearly zero… 1) cost of reading everything fell → you can now watch 100% instead of a sample (every ex

When Does Generating More Help? Disentangling Fixed-Source Synthesis from Source Expansion in Synthetic Data Scaling

ResearchDGX agent

arXiv:2607.01727v1 Announce Type: new Abstract: Synthetic data can be scaled along two routes: Source Expansion (SE), which enlarges the source by adding seed materials or generators, and Fixed-Source

When Sample Selection Bias Precipitates Model Collapse

SafetyDGX agent

arXiv:2606.13732v2 Announce Type: replace Abstract: The proliferation of recursive training on synthetic data can alleviate data scarcity but risks model collapse, where repeated training erodes distr

When Should Service Agents Reconsider? Difficulty-Routed Control in Customer-Service Operations

SafetyDGX agent

arXiv:2607.01426v1 Announce Type: new Abstract: Autonomous customer-service agents are shifting from conversational interfaces toward operational execution roles: they retrieve firm records, apply ser

While it is obviously true that not having verifiable domains makes training models in those spaces difficult... it is also true that models…

ApplicationsDGX agent

While it is obviously true that not having verifiable domains makes training models in those spaces difficult... it is also true that models are also getting much better at non-verifiable domains. The

Who Gets the Reward & Who Gets the Blame? Evaluation-Aligned Training Signals for Multi-LLM Agents

Local AiDGX agent

arXiv:2511.10687v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) in multi-agent systems (MAS) have shown promise for complex tasks, yet current training methods lack principled w

Why Can't I Open My Drawer? Mitigating Object-Driven Shortcuts in Zero-Shot Compositional Action Recognition

TutorialsDGX agent

arXiv:2601.16211v3 Announce Type: replace-cross Abstract: Zero-Shot Compositional Action Recognition (ZS-CAR) requires recognizing novel verb-object combinations composed of previously observed primit

Will Scaling Improve Social Simulation with LLMs?

ResearchDGX agent

arXiv:2607.02464v1 Announce Type: new Abstract: Large Language Model (LLM) social simulations are a promising research method, but they are not yet faithful enough to be adopted widely. In this work,

← Previous
1…382383384385386…1453
Next →