AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,574 results
1 Jul 2026

UniCoder: Unified Visual-to-Code Generation via Symbolic Rewards and Reference-Guided Code Optimization

Model ReleasesDGX agent

arXiv:2606.31732v1 Announce Type: new Abstract: Visual-to-Code generation, which transforms scientific plots, vector graphics, and webpages into executable scripts, demands a level of pixel-precise al

Unified Structural-Hydrodynamic Modeling of Underwater Underactuated Mechanisms and Soft Robots

Model ReleasesDGX agent

arXiv:2603.07939v2 Announce Type: replace Abstract: Underwater robots are widely deployed for ocean exploration and manipulation. Underactuated mechanisms are advantageous in aquatic environments beca

We are at a turning point: many of the decisions we make about AI today will permanently shape our future. Governments and the public need t…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

We are at a turning point: many of the decisions we make about AI today will permanently shape our future. Governments and the public need to clearly understand the impacts, risks, and opportunities o

We are SO back

Model ReleasesDGX agent

We are SO back Claude Fable 5 will be available again globally tomorrow. After a series of productive conversations with the US government, we're redeploying the model with a new set of classifiers to

What Counts as an Error? Dual-Reference Benchmarking for Atypical ASR

Model ReleasesDGX agent

arXiv:2606.31112v1 Announce Type: new Abstract: ASR systems have been often reported to underperform on atypical speech. An often conflated compounding factor is the existence of two valid transcripti

What If We Allocate Test-Time Compute Adaptively?

Model ReleasesDGX agent

arXiv:2602.01070v5 Announce Type: replace Abstract: Test-time compute scaling allocates inference computation uniformly, uses fixed sampling strategies, and applies verification only for reranking. In

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States

Model ReleasesDGX agent

arXiv:2606.31612v1 Announce Type: new Abstract: Mobile GUI agents increasingly face long-horizon tasks that require reading, updating, and reusing task-relevant data across pages and applications. Exi

When Does Learning to Stop Help? A Cost-Aware Study of Early Exits in Reasoning Models

Model ReleasesDGX agent

arXiv:2606.30852v1 Announce Type: new Abstract: Reasoning models spend different amounts of useful computation across instances, but it remains unclear when a learned stopping rule improves over simpl

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

Model ReleasesDGX agent

arXiv:2606.32029v1 Announce Type: cross Abstract: While large language models (LLMs) perform well on table tasks, they still make data referencing errors (DREs), i.e., incorrectly citing or omitting t

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

Model ReleasesDGX agent

arXiv:2606.31307v1 Announce Type: new Abstract: Large language models used in task-oriented dialogue often produce fluent but unsafe responses when backend database calls fail, return empty results, o

Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineering

Model ReleasesDGX agent

arXiv:2606.30911v1 Announce Type: new Abstract: ML engineering agents waste compute rediscovering known techniques because every competition is a cold start. We present HASTE, a hierarchical multi-age

WIDER-FAIR: An Annotated Version of the WIDER-FACE Dataset for Fairness Evaluation

Model ReleasesDGX agent

arXiv:2606.31704v1 Announce Type: new Abstract: The deployment of face detection models in real-world applications raises important fairness concerns, as these systems may showcase performance dispari

Wiki!!!

Model ReleasesDGX agent

Wiki!!! One unexpected outcome of this is that I'm now using the wiki as the ONLY place I run Claude Code I use it as a master controller for all of my repos, kicking off cross-repo tasks and using To

Wisdom Of The (AI) Crowd: Investigating Artificial Swarm Intelligence In Large Language Models

Model ReleasesDGX agent

arXiv:2606.31404v1 Announce Type: new Abstract: Human swarm intelligence demonstrates remarkable collective accuracy but faces scalability constraints in cost, coordination, and time. We investigate w

Wordle 1,837 6/6 ⬛⬛⬛⬛⬛ ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛⬛ ⬛🟩⬛⬛🟩 🟩🟩⬛⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This entry documents a Wordle puzzle solution (puzzle #1,837) shared by Anthropic on X/Twitter, showing the complete sequence of guesses and letter feedback that led to solving the word on the sixth a

Wordle 1,838 4/6 ⬛⬛⬛🟨🟨 ⬛⬛⬛⬛⬛ ⬛⬛🟨⬛🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This appears to be a Wordle game result shared by Anthropic on X (formerly Twitter), showing the solution was found in 4 attempts with a specific pattern of correct (green), present but misplaced (yel

World-Model Collapse as a Phase Transition

Model ReleasesDGX agent

arXiv:2606.31399v1 Announce Type: new Abstract: Water looks unchanged as it warms, then at a critical point it boils. We ask whether long-horizon language agents show an analogous transition in their

WorldRoamBench: An Open-World Benchmark for Long-Horizon Stability of Interactive World Models

Model ReleasesDGX agent

arXiv:2606.31672v1 Announce Type: cross Abstract: Despite rapid progress in interactive world models (IWMs), existing benchmarks evaluate action following only at trajectory level and ignore memory an

Xiaomi-GUI-0 Technical Report

Model ReleasesDGX agent

arXiv:2606.31410v1 Announce Type: new Abstract: Graphical user interface (GUI) agents build on vision-language models to complete user tasks end-to-end in real applications through interface actions s

Yes! Pre-classifying routers are going to result in a lot of bad work because routing is hard and tend to underestimate the value of intelli…

Model ReleasesDGX agent

Yes! Pre-classifying routers are going to result in a lot of bad work because routing is hard and tend to underestimate the value of intelligence on many problems. OpenAI learned this with GPT-5, now

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between…

Model ReleasesDGX agent

You really need to benchmark models for your use case. As soon as judgements & decisions stack on top of each other, the differences between models amplifies, and no standard benchmark will tell you t

Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2606.31846v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language instructions, visual observations, and

30 Jun 2026

3D Scene-Adaptive Trajectory-Controllable Human Image Animation with Camera Movement

Model ReleasesDGX agent

arXiv:2606.30514v1 Announce Type: new Abstract: Human image animation, which aims to generate a video of a reference subject following a provided action sequence, has received increasing research inte

A Bayesian latent Gaussian process framework for aerodynamic uncertainty quantification

Model ReleasesDGX agent

arXiv:2606.28871v1 Announce Type: cross Abstract: Predicting the aerodynamic performance (e.g. lift, drag, and moment coefficients) of an aircraft is challenging -- computational models are biased and

A Comparative Study on Affective Cues in Text Embeddings Across Psychological Emotion Theories

Model ReleasesDGX agent

arXiv:2606.29068v1 Announce Type: cross Abstract: Text encoders are known for their utility in natural language processing, as they are able to efficiently compress inputs into dense vectors while pre

A Diagnostic Framework and Multi-Evaluator Audit of Evaluator-Driven Preference Dynamics in Self-Adapting LLM Agents

Model ReleasesDGX agent

arXiv:2606.29719v1 Announce Type: cross Abstract: Measurements of proprietary LLM evaluators can become invalid within weeks -- we document one case and provide the diagnostic framework to detect it.

A Machine-Verified Proof of a Quantum-Optimization Conjecture

Model ReleasesDGX agent

arXiv:2606.29687v1 Announce Type: cross Abstract: We report a machine-verified resolution of a problem open for over a decade in quantum optimization: the Farhi, Goldstone and Gutmann (FGG) conjecture

A multi-architecture study of specificity refinement and false-positive mechanism analysis in prostate MRI

Model ReleasesDGX agent

arXiv:2606.29977v1 Announce Type: cross Abstract: Objectives: To characterize residual false positives in prostate MRI detection, and to evaluate a lightweight post-hoc refinement head for case-level

A Multi-Dataset Benchmark for Evaluating LLM Agents in Microservice Failure Diagnosis

Model ReleasesDGX agent

arXiv:2606.29193v1 Announce Type: cross Abstract: LLM-based agents are reshaping microservice operations into AgentOps, where benchmarks are key to evaluating failure diagnosis over multimodal observa

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models

Model ReleasesDGX agent

arXiv:2606.28757v1 Announce Type: new Abstract: Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi

A Probabilistic Approach to Trajectory-Based Optimal Experimental Design

Model ReleasesDGX agent

arXiv:2601.11473v2 Announce Type: replace-cross Abstract: We present a novel probabilistic approach for optimal experimental path design. In this approach a discrete path optimization problem is defin

A Stochastic--Geometric Theory of Scaling Laws in Grokking

Model ReleasesDGX agent

arXiv:2606.30388v1 Announce Type: cross Abstract: Delayed generalization (ie~grokking) refers to the phenomenon in which a neural network fits its training data early in training but only begins to ge

Accelerating Q-learning through Efficient Value-Sharing across Actions

Model ReleasesDGX agent

arXiv:2606.29806v1 Announce Type: cross Abstract: Action-values are foundational to many control algorithms such as Q-learning. Therefore learning action-values efficiently is central to reinforcement

Accelerating scientific discovery with Co-Scientist

Model ReleasesDGX agent

arXiv:2502.18864v2 Announce Type: replace Abstract: Scientific discovery is driven by scientists generating novel hypotheses for complex problems that undergo rigorous experimental validation. To augm

Adaptive Financial Transformer with Regime-Gated Attention for Stock Return Prediction

Model ReleasesDGX agent

arXiv:2606.29347v1 Announce Type: cross Abstract: Adaptive Financial Transformer (AFT) is proposed for stock return prediction under non-stationary financial markets. The model incorporates a Market R

ADEPT: An Entropy-Driven Dual-Strategy Agent for Interactive Video Retrieval

Model ReleasesDGX agent

arXiv:2606.28326v1 Announce Type: cross Abstract: This research aims to solve the challenge of video retrieval from massive datasets, caused by ambiguous user queries. Prevailing single-round retrieva

AerialMetric: Benchmarking and Adapting UAV Monocular Metric Depth Estimation in the Real World

Model ReleasesDGX agent

arXiv:2606.29716v1 Announce Type: new Abstract: This paper addresses the problem of monocular metric depth estimation in aerial UAV imagery. Although recent data-driven methods have achieved remarkabl

AERIS: Aerial-Edge Role-Driven Intelligence at Runtime via Orchestrated Language-Model Swarm

Model ReleasesDGX agent

arXiv:2606.30151v1 Announce Type: new Abstract: Integrating large language models into robotic systems holds promise for enhancing autonomy, yet practical deployment remains constrained by strict hear

Agent-Computer Observation Interfaces Enable Dynamic Computer Use

Model ReleasesDGX agent

arXiv:2606.29472v1 Announce Type: new Abstract: SWE-agent established the action interface as an underexplored design axis for software-engineering agents; we make the analogous case for the observati

Agentic Abstention: Do Agents Know When to Stop Instead of Act?

Model ReleasesDGX agent

arXiv:2606.28733v1 Announce Type: new Abstract: LLM agents are expected to act over multiple turns, using search, browsing interfaces, and terminal tools to complete user goals. Yet not every goal is

Agile Reinforcement Learning through Separable Neural Architecture and Applications

Model ReleasesDGX agent

arXiv:2601.23225v2 Announce Type: replace-cross Abstract: Deep reinforcement learning (RL) is increasingly deployed in resource-constrained environments, yet go-to function approximators - multilayer

Alternative Graph Neural Networks: Synergizing GEV Models and Deep Learning for Travel Mode Choice Modeling

Model ReleasesDGX agent

arXiv:2509.07123v2 Announce Type: replace-cross Abstract: Generalized extreme value models capture dependence among choice alternatives in discrete choice modeling, but require this dependence to be p

An Agentic AI Pipeline for Appliance-Level Energy Anomaly Detection and LLM-Driven Recommendations

Model ReleasesDGX agent

arXiv:2606.28467v1 Announce Type: cross Abstract: Appliance-level energy monitoring in office buildings produces noisy alerts that non-expert facility managers struggle to use. This paper proposes an

An AI agent for treatment reasoning over a biomedical tool universe

Model ReleasesDGX agent

arXiv:2606.28692v1 Announce Type: new Abstract: Treatment reasoning underpins every therapeutic decision, integrating disease context, comorbidities, medications, contraindications, and evolving biome

Analysis of Adam Algorithms for Stochastic Dynamic Systems

Model ReleasesDGX agent

arXiv:2606.28879v1 Announce Type: new Abstract: The adaptive moment estimation algorithm, known as Adam, is widely used in modern machine learning, owing to its low per-iteration complexity and strong

Analysis of Parameter Settings for the Bat Algorithm Using Variance Evolution

Model ReleasesDGX agent

arXiv:2606.28644v1 Announce Type: cross Abstract: Parameter settings in evolutionary algorithms and metaheuristics are important because such parameter values can influence the performance of algorith

Animation2Code: Evaluating Temporal Visual Reasoning in Video-to-Code Generation

Model ReleasesDGX agent

arXiv:2606.28593v1 Announce Type: cross Abstract: While recent vision-language models (VLMs) have achieved significant improvements on static visual-to-code tasks such as generating code for webpages,

Anisotropy Decides Cosine vs. Rank Metrics for Text Embeddings

Model ReleasesDGX agent

arXiv:2606.29571v1 Announce Type: new Abstract: The standard way to compare two text embeddings is cosine similarity. Scattered studies report that a different metric does better, but never pin down t

Anthropic integration with Modal brings scalable compute to Claude Science

Model ReleasesDGX agent

Anthropic has integrated Claude with Modal's serverless compute platform, enabling scalable cloud infrastructure for scientific computing workflows. This integration allows Claude to leverage Modal's

Anthropic launches Claude Science, an AI workbench that uses existing Claude models like Opus 4.8 to integrate 60+ scientific databases and specialized toolkits (Rebecca Bellan/TechCrunch)

Model ReleasesDGX agent

Rebecca Bellan / TechCrunch: Anthropic launches Claude Science, an AI workbench that uses existing Claude models like Opus 4.8 to integrate 60+ scientific databases and specialized toolkits — Anthropi

Anthropic launches Claude Sonnet 5, saying it nears Opus 4.8 performance at lower prices and is substantially better than Sonnet 4.6 for agentic work (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic launches Claude Sonnet 5, saying it nears Opus 4.8 performance at lower prices and is substantially better than Sonnet 4.6 for agentic work — Claude Sonnet 5 is built to be the mo

Anthropic says the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5 and that it will begin restoring access Wednesday (@anthropicai)

Model ReleasesDGX agent

@anthropicai: Anthropic says the Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5 and that it will begin restoring access Wednesday — We've received notice that the Dep

Anthropic’s long-sidelined Fable 5 is greenlit to return

Model ReleasesDGX agent

After weeks of negotiating with the Trump administration, Anthropic is finally going to be able to bring Claude Fable 5 back online. In a post on X, Anthropic said it plans to begin restoring access W

Anti-Collapse Dynamics and the Emergence of Multi-Time-Scale Learning in Recurrent Neural Networks

Model ReleasesDGX agent

arXiv:2606.29519v1 Announce Type: new Abstract: Long-range learning is hard for recurrent networks trained with stochastic gradient descent, because the influence of a past input fades with the lag el

Argus: Metric Panoramic 3D Reconstruction for Indoor Scenes

Model ReleasesDGX agent

arXiv:2606.30047v1 Announce Type: new Abstract: Metric feed-forward 3D reconstruction for panoramic data remains under-explored due to the lack of large-scale panoramic RGB-D training data. We present

Arko-T: A Foundation Model for Text-to-Structured 3D Generation

Model ReleasesDGX agent

arXiv:2606.30429v1 Announce Type: new Abstract: Text-to-3D systems can now synthesize a mechanical part from a single sentence, yet the result is a shape to render, not a design to edit. We present Ar

ARMOR: Adaptive Retriever Optimization for Low-Resource Telecom Question Answering

Model ReleasesDGX agent

arXiv:2606.29706v1 Announce Type: cross Abstract: Telecom question answering (QA) is a challenging setting for retrieval-augmented generation (RAG): evidence is fragmented across standards, papers, en

Assessing the Business Process Modeling Competences of Large Language Models

Model ReleasesDGX agent

arXiv:2601.21787v2 Announce Type: replace-cross Abstract: The creation of Business Process Model and Notation (BPMN) models is a complex and time-consuming task requiring both domain knowledge and pro

Attractor States Emerge in Multi-Turn LLM Conversations

Model ReleasesDGX agent

arXiv:2606.30571v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in open-ended multi-agent settings, but the long-run dynamics of model--model interaction remain po

Attribution Graphs and Causal Probing for Mechanistic Discovery and Bias Repair in Multimodal Generative Learning

Model ReleasesDGX agent

arXiv:2510.12957v4 Announce Type: replace-cross Abstract: We treat the internals of generative models as mechanistic objects rather than black boxes. We introduce extbf{Attribution Graphs} (AGs), whic

← Previous
1…116117118119120…377
Next →