AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,456 results
Model Releases

What If We Allocate Test-Time Compute Adaptively?

DGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

model-releasesarxiv-cs-cl
1 Jul 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

What is your definition of a forward deployed engineer? asks @latentspacepod. @natalie_meurer: 'That is really the point of my session: the …

DGX agent

What is your definition of a forward deployed engineer? asks @latentspacepod. @natalie_meurer: 'That is really the point of my session: the role lacks a consistent definition. If you look at its histo

agentsswyx--x
1 Jul 2026
Model Releases

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States

DGX agent

arXiv:2606.31612v1 Announce Type: new Abstract: Mobile GUI agents increasingly face long-horizon tasks that require reading, updating, and reusing task-relevant data across pages and applications. Exi

model-releasesarxiv-cs-cv
1 Jul 2026
Safety

What Probing Reveals about Autonomous Driving: Linking Internal Prediction Errors to Ego Planning

DGX agent

arXiv:2606.31106v1 Announce Type: cross Abstract: Large-scale datasets and fast simulators have enabled improvements in driving policies that appear safe and robust, yet strong performance in nominal

safetyarxiv-cs-ai
1 Jul 2026
Hardware

When boomer companies get high Anthropic bill, they set spend limits When I see a nearly million dollar Anthropic monthly bill, my first rea…

DGX agent

When boomer companies get high Anthropic bill, they set spend limits When I see a nearly million dollar Anthropic monthly bill, my first reaction is to complain about use of shitty Haiku models Imagin

hardwaredylan-patel--x
1 Jul 2026
Research

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs

DGX agent

arXiv:2606.30814v1 Announce Type: new Abstract: Calibration evaluates whether a model confidence aligns with its empirical accuracy. Existing studies often compare the calibration of different large l

researcharxiv-cs-cl
1 Jul 2026
Model Releases

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

DGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

model-releasesarxiv-cs-ai
1 Jul 2026
Research

When few labeled target data suffice: a theory of semi-supervised domain adaptation via fine-tuning from multiple adaptive starts

DGX agent

arXiv:2507.14661v2 Announce Type: replace-cross Abstract: Semi-supervised domain adaptation (SSDA) seeks to achieve accurate predictions in a target domain with limited labeled target data by exploiti

researcharxiv-cs-lg
1 Jul 2026
Model Releases

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

DGX agent

arXiv:2606.32029v1 Announce Type: cross Abstract: While large language models (LLMs) perform well on table tasks, they still make data referencing errors (DREs), i.e., incorrectly citing or omitting t

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

When Regulation Has Memory: Hysteresis and Control Burden in Artificial Agency

DGX agent

arXiv:2606.30975v1 Announce Type: new Abstract: Adaptive agents are usually judged by what they do, but an agent can appear stable while the internal effort required to keep it stable is increasing. T

agentsarxiv-cs-ai
1 Jul 2026
Research

When Reranking Hurts: Uncertainty-Based Gating for Few-Shot Reranking

DGX agent

arXiv:2606.31087v1 Announce Type: cross Abstract: Few-shot selection typically assumes that reranking retrieved examples always improves performance. We challenge this view by identifying that the exp

researcharxiv-cs-ai
1 Jul 2026
Research

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models

DGX agent

arXiv:2604.03316v2 Announce Type: replace Abstract: Attention sinks are defined as tokens that attract disproportionate attention. While these have been studied in single modality transformers, their

researcharxiv-cs-cv
1 Jul 2026
Model Releases

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

DGX agent

arXiv:2606.31307v1 Announce Type: new Abstract: Large language models used in task-oriented dialogue often produce fluent but unsafe responses when backend database calls fail, return empty results, o

model-releasesarxiv-cs-cl
1 Jul 2026
Research

When to Truncate a Feature Ranking: A Residual-Overlap Stopping Rule for Subset Selection

DGX agent

arXiv:2606.31686v1 Announce Type: cross Abstract: Feature rankings are widely used in supervised feature selection because they are simple, scalable and easy to interpret. Variables are first ranked b

researcharxiv-cs-ai
1 Jul 2026
Local Ai

When transformers learn 'impossible' languages, what do they learn?

DGX agent

arXiv:2606.30815v1 Announce Type: cross Abstract: Recent work suggests that transformer language models show a bias towards human languages over unnatural ('impossible') languages argued to be unacqui

local-aiarxiv-cs-ai
1 Jul 2026
Industry

When you can cheat on every test, grab quotes from every book, delegate heavy workloads, then why would any human even get out of bed? For p…

DGX agent

When you can cheat on every test, grab quotes from every book, delegate heavy workloads, then why would any human even get out of bed? For purpose. For the challenge. For the thrill of growth. For sel

industryallie-k--miller--x
1 Jul 2026
Agents

when you have fomo but not enough use cases

DGX agent

This post likely discusses the paradox of experiencing fear of missing out (FOMO) on emerging technologies or trends, while lacking practical applications or use cases to justify adoption. The content

agentsjerry-liu--x
1 Jul 2026
Local Ai

Which Tokens Matter? Adaptive Token Selection for RLVR with the Relative Surprisal Index

DGX agent

arXiv:2606.31575v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a powerful tool for propelling Large Language Models (LLMs) beyond imitation-based training towards more robust r

local-aiarxiv-cs-ai
1 Jul 2026
Research

While working Americans are struggling to make ends meet, Trump enriched himself to the tune of $2.2 billion last year. A foreign government…

DGX agent

While working Americans are struggling to make ends meet, Trump enriched himself to the tune of $2.2 billion last year. A foreign government bought into his crypto business and then got a sweetheart d

researchyann-lecun--x
1 Jul 2026
Research

Who Determines the Meaning of an Emotion? Affective Sovereignty as an Epistemic Consequence of Measurement Limits

DGX agent

arXiv:2606.31442v1 Announce Type: new Abstract: Emotion-sensing AI is rapidly becoming embedded in vehicles, home appliances, dialogue agents, and social infrastructure, giving rise to a sphere in whi

researcharxiv-cs-ai
1 Jul 2026
Research

Who did it best? GLM-5.2 (left) | Fugu Ultra (middle) | Fable 5 (right) Same one-shot prompt. The last one is my favorite!

DGX agent

This post compares the outputs of three AI models—GLM-5.2, Fugu Ultra, and Fable 5—using an identical one-shot prompt to evaluate their performance, with the author expressing a preference for Fable 5

researchdair-ai--x
1 Jul 2026
Research

Why Do Few-Step Text Latents Fail When Image Latents Work? Non-Commitment at Sharp Categorical Readouts

DGX agent

arXiv:2606.30705v1 Announce Type: cross Abstract: Deterministic few-step generation succeeds on continuous image latents but collapses to incoherent text on continuous text latents, and we show the ca

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineering

DGX agent

arXiv:2606.30911v1 Announce Type: new Abstract: ML engineering agents waste compute rediscovering known techniques because every competition is a cold start. We present HASTE, a hierarchical multi-age

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

WIDER-FAIR: An Annotated Version of the WIDER-FACE Dataset for Fairness Evaluation

DGX agent

arXiv:2606.31704v1 Announce Type: new Abstract: The deployment of face detection models in real-world applications raises important fairness concerns, as these systems may showcase performance dispari

model-releasesarxiv-cs-cv
1 Jul 2026
Model Releases

Wiki!!!

DGX agent

Wiki!!! One unexpected outcome of this is that I'm now using the wiki as the ONLY place I run Claude Code I use it as a master controller for all of my repos, kicking off cross-repo tasks and using To

model-releasesharrison-chase--x
1 Jul 2026
Agents

wikis for memory are all the rage - we wrote about this yesterday today we're releasing an open source example for doing this code bases

DGX agent

This post announces the release of an open source tool or example demonstrating how to implement wikis for memory management, following up on a previous discussion about this trending approach. The re

agentsharrison-chase--x
1 Jul 2026
Research

WildProp: Visual Estimation of Wildlife Body Proportions at Scale

DGX agent

arXiv:2606.31125v1 Announce Type: new Abstract: Population-level morphometric measurements underpin ecological and evolutionary studies but traditionally require controlled imaging or physical specime

researcharxiv-cs-cv
1 Jul 2026
Safety

Will the AI bubble pop is the wrong question. We partnered with Damon Cassidy, a video essayist with ~300k subscribers: even if it pops, the…

DGX agent

Will the AI bubble pop is the wrong question. We partnered with Damon Cassidy, a video essayist with ~300k subscribers: even if it pops, the race to uncontrollable superintelligence doesn't go away. P

safetyconnor-leahy--x
1 Jul 2026
Research

Wind and State Estimation on SE(3): Comparative Evaluation of EKF and UKF with Continuous and Discrete Quadrotor Models

DGX agent

arXiv:2606.30804v1 Announce Type: new Abstract: Use of quadrotor UAVs for wind velocity estimation is gaining popularity in recent studies, leveraging their maneuverability, compact size and low cost.

researcharxiv-cs-ro
1 Jul 2026
Model Releases

Wisdom Of The (AI) Crowd: Investigating Artificial Swarm Intelligence In Large Language Models

DGX agent

arXiv:2606.31404v1 Announce Type: new Abstract: Human swarm intelligence demonstrates remarkable collective accuracy but faces scalability constraints in cost, coordination, and time. We investigate w

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Wordle 1,837 6/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛⬛ ⬛🟩⬛⬛🟩 🟩🟩⬛⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This entry documents a Wordle puzzle solution (puzzle #1,837) shared by Anthropic on X/Twitter, showing the complete sequence of guesses and letter feedback that led to solving the word on the sixth a

model-releasesanthropic--x
1 Jul 2026
Model Releases

Wordle 1,838 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛🟩 🟩🟩🟩🟩🟩

DGX agent

This appears to be a Wordle game result shared by Anthropic on X (formerly Twitter), showing the solution was found in 4 attempts with a specific pattern of correct (green), present but misplaced (yel

model-releasesanthropic--x
1 Jul 2026
Model Releases

World-Model Collapse as a Phase Transition

DGX agent

arXiv:2606.31399v1 Announce Type: new Abstract: Water looks unchanged as it warms, then at a critical point it boils. We ask whether long-horizon language agents show an analogous transition in their

model-releasesarxiv-cs-ai
1 Jul 2026
Tutorials

World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration

DGX agent

arXiv:2606.31946v1 Announce Type: new Abstract: The fundamental obstacle to industrial grade video generation is the lack of controllability: existing models treat video as a pixel distribution sampli

tutorialsarxiv-cs-cv
1 Jul 2026
Model Releases

WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World Models

DGX agent

arXiv:2606.31672v1 Announce Type: cross Abstract: Despite rapid progress in interactive world models (IWMs), existing benchmarks evaluate action following only at trajectory level and ignore memory an

model-releasesarxiv-cs-ai
1 Jul 2026
Tools

Wow and soon after that @mitsuhiko came by and also expressed interest in analysis of token usage in reviews, a perfect use case for context…

DGX agent

Wow and soon after that @mitsuhiko came by and also expressed interest in analysis of token usage in reviews, a perfect use case for context-lens! I didn’t wake up thinking I’d show my work to @simonw

toolsswyx--x
1 Jul 2026
Model Releases

Xiaomi-GUI-0 Technical Report

DGX agent

arXiv:2606.31410v1 Announce Type: new Abstract: Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions s

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Yes! Pre-classifying routers are going to result in a lot of bad work because routing is hard and tend to underestimate the value of intelli…

DGX agent

Yes! Pre-classifying routers are going to result in a lot of bad work because routing is hard and tend to underestimate the value of intelligence on many problems. OpenAI learned this with GPT-5, now

model-releasesethan-mollick--x
1 Jul 2026
Tools

yesterday Fable re-release got announced and today we’re hearing from Thariq @trq212

DGX agent

Fable's re-release was recently announced, and Thariq (Twitter handle @trq212) provided commentary or insights on the news in a follow-up discussion. The post appears to be from Swyx sharing reactions

toolsswyx--x
1 Jul 2026
Agents

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, …

DGX agent

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, proposed by @a1zhang from MIT, help: instead of working mode

agentsharrison-chase--x
1 Jul 2026
Agents

You can now run recursive language model (RLM) workflows in Deep Agents. Everything you need to know in 6 minutes from @sydneyrunkle.

DGX agent

Recursive Language Model (RLM) workflows are now available as a feature in Deep Agents, allowing AI systems to iteratively call language models within multi-step processes. This capability enables mor

agentsharrison-chase--x
1 Jul 2026
Agents

You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3️⃣ Enter your API key …and you're ready …

DGX agent

You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3️⃣ Enter your API key …and you're ready to go with frontier performance on open weights. Great demo

agentsharrison-chase--x
1 Jul 2026
Model Releases

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

DGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

model-releasesethan-mollick--x
1 Jul 2026
Agents

Your site, your rules: new AI traffic options for all customers

DGX agent

For our second Content Independence Day, we’re giving website owners finer options to manage AI traffic. Instead of a one-size-fits-all block, all customers can now easily distinguish and manage Searc

agentscloudflare-ai
1 Jul 2026
Model Releases

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

DGX agent

arXiv:2606.31846v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language instructions, visual observations, and

model-releasesarxiv-cs-ai
1 Jul 2026
Research

ZEBRA: Zero-Shot Entropy-Regularized Prompt Learning for Base-to-Novel Generalization in Audio-Language Models

DGX agent

arXiv:2606.31587v1 Announce Type: cross Abstract: Audio-Language Models (ALMs) achieve strong zero-shot performance by aligning audio with textual class descriptions. Although prompt learning improves

researcharxiv-cs-ai
1 Jul 2026
Research

Zero-Shot Quantization for Object Detectors using Off-the-Shelf Generative Models

DGX agent

arXiv:2606.31456v1 Announce Type: new Abstract: With an increasing number of Object Detection (OD) models being deployed on edge devices, Zero-Shot Quantization for OD (ZSQ-OD) aims to quantize these

researcharxiv-cs-lg
1 Jul 2026
Industry

1/ Today, more than 140 companies, most of which compete fiercely with one another, agreed to back the same stablecoin. The vehicle is @open…

DGX agent

1/ Today, more than 140 companies, most of which compete fiercely with one another, agreed to back the same stablecoin. The vehicle is @openstandard, a new and deliberately independent company launchi

industryemad-mostaque--x
30 Jun 2026
← Previous
1…545546547548549…1843
Next →