AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,343
  • Agents7,552
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,164
  • Local Ai4,928
  • Model Releases23,861
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,402

Source
Human
88,343Total entries
1Added by human
88,342Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
88,342 results
1 Jul 2026

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

Model ReleasesDGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

When few labeled target data suffice: a theory of semi-supervised domain adaptation via fine-tuning from multiple adaptive starts

ResearchDGX agent

arXiv:2507.14661v2 Announce Type: replace-cross Abstract: Semi-supervised domain adaptation (SSDA) seeks to achieve accurate predictions in a target domain with limited labeled target data by exploiti

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.32029v1 Announce Type: cross Abstract: While large language models (LLMs) perform well on table tasks, they still make data referencing errors (DREs), i.e., incorrectly citing or omitting t

When Regulation Has Memory: Hysteresis and Control Burden in Artificial Agency

AgentsDGX agent

arXiv:2606.30975v1 Announce Type: new Abstract: Adaptive agents are usually judged by what they do, but an agent can appear stable while the internal effort required to keep it stable is increasing. T

When Reranking Hurts: Uncertainty-Based Gating for Few-Shot Reranking

ResearchDGX agent

arXiv:2606.31087v1 Announce Type: cross Abstract: Few-shot selection typically assumes that reranking retrieved examples always improves performance. We challenge this view by identifying that the exp

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models

ResearchDGX agent

arXiv:2604.03316v2 Announce Type: replace Abstract: Attention sinks are defined as tokens that attract disproportionate attention. While these have been studied in single modality transformers, their

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

Model ReleasesDGX agent

arXiv:2606.31307v1 Announce Type: new Abstract: Large language models used in task-oriented dialogue often produce fluent but unsafe responses when backend database calls fail, return empty results, o

When to Truncate a Feature Ranking: A Residual-Overlap Stopping Rule for Subset Selection

ResearchDGX agent

arXiv:2606.31686v1 Announce Type: cross Abstract: Feature rankings are widely used in supervised feature selection because they are simple, scalable and easy to interpret. Variables are first ranked b

When transformers learn 'impossible' languages, what do they learn?

Local AiDGX agent

arXiv:2606.30815v1 Announce Type: cross Abstract: Recent work suggests that transformer language models show a bias towards human languages over unnatural ('impossible') languages argued to be unacqui

When you can cheat on every test, grab quotes from every book, delegate heavy workloads, then why would any human even get out of bed? For p…

IndustryDGX agent

When you can cheat on every test, grab quotes from every book, delegate heavy workloads, then why would any human even get out of bed? For purpose. For the challenge. For the thrill of growth. For sel

when you have fomo but not enough use cases

AgentsDGX agent

This post likely discusses the paradox of experiencing fear of missing out (FOMO) on emerging technologies or trends, while lacking practical applications or use cases to justify adoption. The content

Which Tokens Matter? Adaptive Token Selection for RLVR with the Relative Surprisal Index

Local AiDGX agent

arXiv:2606.31575v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a powerful tool for propelling Large Language Models (LLMs) beyond imitation-based training towards more robust r

While working Americans are struggling to make ends meet, Trump enriched himself to the tune of $2.2 billion last year. A foreign government…

ResearchDGX agent

While working Americans are struggling to make ends meet, Trump enriched himself to the tune of $2.2 billion last year. A foreign government bought into his crypto business and then got a sweetheart d

Who Determines the Meaning of an Emotion? Affective Sovereignty as an Epistemic Consequence of Measurement Limits

ResearchDGX agent

arXiv:2606.31442v1 Announce Type: new Abstract: Emotion-sensing AI is rapidly becoming embedded in vehicles, home appliances, dialogue agents, and social infrastructure, giving rise to a sphere in whi

Who did it best? GLM-5.2 (left) | Fugu Ultra (middle) | Fable 5 (right) Same one-shot prompt. The last one is my favorite!

ResearchDGX agent

This post compares the outputs of three AI models—GLM-5.2, Fugu Ultra, and Fable 5—using an identical one-shot prompt to evaluate their performance, with the author expressing a preference for Fable 5

Why Do Few-Step Text Latents Fail When Image Latents Work? Non-Commitment at Sharp Categorical Readouts

ResearchDGX agent

arXiv:2606.30705v1 Announce Type: cross Abstract: Deterministic few-step generation succeeds on continuous image latents but collapses to incoherent text on continuous text latents, and we show the ca

Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineering

Model ReleasesDGX agent

arXiv:2606.30911v1 Announce Type: new Abstract: ML engineering agents waste compute rediscovering known techniques because every competition is a cold start. We present HASTE, a hierarchical multi-age

WIDER-FAIR: An Annotated Version of the WIDER-FACE Dataset for Fairness Evaluation

Model ReleasesDGX agent

arXiv:2606.31704v1 Announce Type: new Abstract: The deployment of face detection models in real-world applications raises important fairness concerns, as these systems may showcase performance dispari

Wiki!!!

Model ReleasesDGX agent

Wiki!!! One unexpected outcome of this is that I'm now using the wiki as the ONLY place I run Claude Code I use it as a master controller for all of my repos, kicking off cross-repo tasks and using To

wikis for memory are all the rage - we wrote about this yesterday today we're releasing an open source example for doing this code bases

AgentsDGX agent

This post announces the release of an open source tool or example demonstrating how to implement wikis for memory management, following up on a previous discussion about this trending approach. The re

WildProp: Visual Estimation of Wildlife Body Proportions at Scale

ResearchDGX agent

arXiv:2606.31125v1 Announce Type: new Abstract: Population-level morphometric measurements underpin ecological and evolutionary studies but traditionally require controlled imaging or physical specime

Will the AI bubble pop is the wrong question. We partnered with Damon Cassidy, a video essayist with ~300k subscribers: even if it pops, the…

SafetyDGX agent

Will the AI bubble pop is the wrong question. We partnered with Damon Cassidy, a video essayist with ~300k subscribers: even if it pops, the race to uncontrollable superintelligence doesn't go away. P

Wind and State Estimation on SE(3): Comparative Evaluation of EKF and UKF with Continuous and Discrete Quadrotor Models

ResearchDGX agent

arXiv:2606.30804v1 Announce Type: new Abstract: Use of quadrotor UAVs for wind velocity estimation is gaining popularity in recent studies, leveraging their maneuverability, compact size and low cost.

Wisdom Of The (AI) Crowd: Investigating Artificial Swarm Intelligence In Large Language Models

Model ReleasesDGX agent

arXiv:2606.31404v1 Announce Type: new Abstract: Human swarm intelligence demonstrates remarkable collective accuracy but faces scalability constraints in cost, coordination, and time. We investigate w

Wordle 1,837 6/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛⬛ ⬛🟩⬛⬛🟩 🟩🟩⬛⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This entry documents a Wordle puzzle solution (puzzle #1,837) shared by Anthropic on X/Twitter, showing the complete sequence of guesses and letter feedback that led to solving the word on the sixth a

Wordle 1,838 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This appears to be a Wordle game result shared by Anthropic on X (formerly Twitter), showing the solution was found in 4 attempts with a specific pattern of correct (green), present but misplaced (yel

World-Model Collapse as a Phase Transition

Model ReleasesDGX agent

arXiv:2606.31399v1 Announce Type: new Abstract: Water looks unchanged as it warms, then at a critical point it boils. We ask whether long-horizon language agents show an analogous transition in their

World Narrative Model for Highly Controllable Video Generation: A Paradigm Shift from Pixel Sampling to Physical World Orchestration

TutorialsDGX agent

arXiv:2606.31946v1 Announce Type: new Abstract: The fundamental obstacle to industrial grade video generation is the lack of controllability: existing models treat video as a pixel distribution sampli

WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World Models

Model ReleasesDGX agent

arXiv:2606.31672v1 Announce Type: cross Abstract: Despite rapid progress in interactive world models (IWMs), existing benchmarks evaluate action following only at trajectory level and ignore memory an

Wow and soon after that @mitsuhiko came by and also expressed interest in analysis of token usage in reviews, a perfect use case for context…

ToolsDGX agent

Wow and soon after that @mitsuhiko came by and also expressed interest in analysis of token usage in reviews, a perfect use case for context-lens! I didn’t wake up thinking I’d show my work to @simonw

Xiaomi-GUI-0 Technical Report

Model ReleasesDGX agent

arXiv:2606.31410v1 Announce Type: new Abstract: Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions s

Yes! Pre-classifying routers are going to result in a lot of bad work because routing is hard and tend to underestimate the value of intelli…

Model ReleasesDGX agent

Yes! Pre-classifying routers are going to result in a lot of bad work because routing is hard and tend to underestimate the value of intelligence on many problems. OpenAI learned this with GPT-5, now

yesterday Fable re-release got announced and today we’re hearing from Thariq @trq212

ToolsDGX agent

Fable's re-release was recently announced, and Thariq (Twitter handle @trq212) provided commentary or insights on the news in a follow-up discussion. The post appears to be from Swyx sharing reactions

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, …

AgentsDGX agent

you can now build RLMs with Deep Agents! the more context agents accumulate, the worse they perform, a phenomenon called context rot. RLMs, proposed by @a1zhang from MIT, help: instead of working mode

You can now run recursive language model (RLM) workflows in Deep Agents. Everything you need to know in 6 minutes from @sydneyrunkle.

AgentsDGX agent

Recursive Language Model (RLM) workflows are now available as a feature in Deep Agents, allowing AI systems to iteratively call language models within multi-step processes. This capability enables mor

You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3️⃣ Enter your API key …and you're ready …

AgentsDGX agent

You can start building with GLM 5.2 in minutes: 1️⃣ Download dcode 2️⃣ Select your model (GLM 5.2) 3️⃣ Enter your API key …and you're ready to go with frontier performance on open weights. Great demo

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

Model ReleasesDGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

Your site, your rules: new AI traffic options for all customers

AgentsDGX agent

For our second Content Independence Day, we’re giving website owners finer options to manage AI traffic. Instead of a one-size-fits-all block, all customers can now easily distinguish and manage Searc

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.31846v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language instructions, visual observations, and

ZEBRA: Zero-Shot Entropy-Regularized Prompt Learning for Base-to-Novel Generalization in Audio-Language Models

ResearchDGX agent

arXiv:2606.31587v1 Announce Type: cross Abstract: Audio-Language Models (ALMs) achieve strong zero-shot performance by aligning audio with textual class descriptions. Although prompt learning improves

Zero-Shot Quantization for Object Detectors using Off-the-Shelf Generative Models

ResearchDGX agent

arXiv:2606.31456v1 Announce Type: new Abstract: With an increasing number of Object Detection (OD) models being deployed on edge devices, Zero-Shot Quantization for OD (ZSQ-OD) aims to quantize these

30 Jun 2026

1/ Today, more than 140 companies, most of which compete fiercely with one another, agreed to back the same stablecoin. The vehicle is @open…

IndustryDGX agent

1/ Today, more than 140 companies, most of which compete fiercely with one another, agreed to back the same stablecoin. The vehicle is @openstandard, a new and deliberately independent company launchi

15 minutes before @swyx is on the main stage at @aiDotEngineer The amount of people visiting the worlds fair this year is unbelievable 😲

ToolsDGX agent

This post captures excitement about an upcoming presentation by Swyx at an AI Engineer conference, noting the unexpectedly large attendance at the event's world fair component. The message was posted

3D Field of Junctions: A Noise-Robust, Training-Free Structural Prior for Volumetric Inverse Problems

ResearchDGX agent

arXiv:2603.02149v2 Announce Type: replace Abstract: Volume denoising is a foundational problem in computational imaging, as many 3D imaging inverse problems face high levels of measurement noise. Insp

3D Scene-Adaptive Trajectory-Controllable Human Image Animation with Camera Movement

Model ReleasesDGX agent

arXiv:2606.30514v1 Announce Type: new Abstract: Human image animation, which aims to generate a video of a reference subject following a provided action sequence, has received increasing research inte

5ting at SemEval-2026 Task 8: Strong End-to-End Multi-Turn RAG via LLM-Based Reranking and Faithfulness Control

ResearchDGX agent

arXiv:2606.28737v1 Announce Type: cross Abstract: We introduce 5ting, our system for the SemEval2026 Task 8 (MTRAGEval), which evaluates multi-turn Retrieval Augmented Generation (RAG) systems. Multi

A Bayesian latent Gaussian process framework for aerodynamic uncertainty quantification

Model ReleasesDGX agent

arXiv:2606.28871v1 Announce Type: cross Abstract: Predicting the aerodynamic performance (e.g. lift, drag, and moment coefficients) of an aircraft is challenging -- computational models are biased and

A binarized-domains arc-consistency algorithm for TCSPs: its computational analysis and its use as a filtering procedure in solution search algorithms

TutorialsDGX agent

arXiv:2002.11508v3 Announce Type: replace Abstract: TCSPs (Temporal Constraint Satisfaction Problems) [Dechter et al. 1991] get rid of unary constraints by binarizing them after having added an 'origi

A candid hallway track interview with the man himself - @swyx Mini Podcast if you will, he's told me about a huge milestone tomorrow, things…

ToolsDGX agent

A candid hallway track interview with the man himself - @swyx Mini Podcast if you will, he's told me about a huge milestone tomorrow, things that have surprised him about @aiDotEngineer and ... a few

A causal modeling perspective on decision theory

SafetyDGX agent

arXiv:2606.29911v1 Announce Type: new Abstract: Decision theory provides a formal framework for how agents should make choices under uncertainty, drawing on ideas from philosophy, probability, and cau

A Classifier-Agnostic Zero-Shot Adversarial Attack Detection via CLIP

ResearchDGX agent

arXiv:2606.30342v1 Announce Type: new Abstract: Adversarial attacks pose a challenge to the reliability of deep learning models, motivating effective detection methods. Existing techniques often rely

A Cognition-Emotion-Personality Framework for Modeling Human-Like Awareness and Behavior in Emergency Evacuations

AgentsDGX agent

arXiv:2606.29212v1 Announce Type: new Abstract: Agent-based evacuation simulations are widely used to study crowd behavior during emergencies, but many models rely on assumptions such as perfect event

a common pattern for (a part of) memory that we see: Wiki Memory Examples: - DeepWiki (@cognition) - AutoWiki (@FactoryAI) - LLM Wiki (@karp…

AgentsDGX agent

This post describes a common architectural pattern for integrating memory components into AI systems, illustrated through examples like DeepWiki and AutoWiki that automatically generate and maintain k

A Comparative Study on Affective Cues in Text Embeddings Across Psychological Emotion Theories

Model ReleasesDGX agent

arXiv:2606.29068v1 Announce Type: cross Abstract: Text encoders are known for their utility in natural language processing, as they are able to efficiently compress inputs into dense vectors while pre

A Conditional GAN for Tabular Data Generation with Probabilistic Sampling of Latent Subspaces

ResearchDGX agent

arXiv:2508.00472v2 Announce Type: replace Abstract: The tabular form constitutes the standard way of representing data in relational database systems and spreadsheets. But, similarly to other forms, t

A conversation is all it takes to launch powerful ComfyUI workflows. This showcase demonstrates how Comfy MCP can generate, refine, and iter…

AgentsDGX agent

A conversation is all it takes to launch powerful ComfyUI workflows. This showcase demonstrates how Comfy MCP can generate, refine, and iterate on creative assets using natural language. To try this w

A Deep Multiscale Neural Network for Accurate Neurological Disorder Detection from MRI Scans and Real-Time Web Deployment

ResearchDGX agent

arXiv:2606.29106v1 Announce Type: cross Abstract: Neurological disorders involve diverse pathologies of the brain and nervous system, making early and accurate detection essential. While many deep CNN

A Deterministic Sampling Method via Maximum Mean Discrepancy Flow with Adaptive Kernel

ResearchDGX agent

arXiv:2111.10722v4 Announce Type: replace-cross Abstract: We propose a novel deterministic sampling method, EVI-MMD, to approximate a target distribution rho^* by minimizing the kernel discrepancy, al

A Diagnostic Framework and Multi-Evaluator Audit of Evaluator-Driven Preference Dynamics in Self-Adapting LLM Agents

Model ReleasesDGX agent

arXiv:2606.29719v1 Announce Type: cross Abstract: Measurements of proprietary LLM evaluators can become invalid within weeks -- we document one case and provide the diagnostic framework to detect it.

A Distributionally Robust Framework for Learned Reconstructions in Inverse Problems

TutorialsDGX agent

arXiv:2606.30230v1 Announce Type: cross Abstract: Learned reconstruction operators for inverse problems are typically trained under a fixed noise model, and generalize poorly when the distribution dur

← Previous
1…434435436437438…1473
Next →