AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
22,620 results
Model Releases

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measur…

DGX agent

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measuring every biomarker, or @sytses openly sharing and analyzing

model-releasesclem-delangue--x
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

The Growing Pains of Frontier Models: When Leaderboards Stop Separating and What to Measure Next

DGX agent

arXiv:2605.18840v1 Announce Type: cross Abstract: Leaderboards rank frontier models on independent axes but do not reveal whether capabilities reinforce or trade off across releases -- and at the fron

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular,…

DGX agent

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular, and represents an important milestone for the math and AI c

model-releasesopenai--x
20 May 2026
Model Releases

The Silent Hyperparameter: Quantifying the Impact of Inference Backends on LLM Reproducibility

DGX agent

arXiv:2605.19537v1 Announce Type: new Abstract: Progress in LLMs is increasingly measured through standardized benchmarks, where state-of-the-art improvements are often separated by fractions of a per

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

The World Won't Stay Still: Programmable Evolution for Agent Benchmarks

DGX agent

arXiv:2603.05910v2 Announce Type: replace Abstract: LLM-powered tool-calling agents fulfill user requests by interacting with environments, querying data, and invoking tools in a multi-turn process. Y

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Theory-optimal Quantization Based on Flatness

DGX agent

arXiv:2605.18800v1 Announce Type: cross Abstract: Post-training quantization has emerged as a widely adopted technique for compressing and accelerating the inference of Large Language Models (LLMs). T

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

This result points to something larger: AI systems are becoming capable of holding together long, difficult chains of reasoning, connecting …

DGX agent

This result points to something larger: AI systems are becoming capable of holding together long, difficult chains of reasoning, connecting ideas across distant fields, and surfacing paths researchers

model-releasesopenai--x
20 May 2026
Model Releases

TideGS: Scalable Training of Over One Billion 3D Gaussian Splatting Primitives via Out-of-Core Optimization

DGX agent

arXiv:2605.20150v1 Announce Type: new Abstract: Training 3D Gaussian Splatting (3DGS) at billion-primitive scale is fundamentally memory-bound: each Gaussian primitive carries a large attribute vector

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Time-optimal neural feedback control of nilpotent systems as a binary classification problem

DGX agent

arXiv:2503.17581v2 Announce Type: replace-cross Abstract: A computational method for the synthesis of time-optimal feedback control laws for linear nilpotent systems is proposed. The method is based o

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

Time to REFLECT: Can We Trust LLM Judges for Evidence-based Research Agents?

DGX agent

arXiv:2605.19196v1 Announce Type: new Abstract: Deep research agents increasingly automate complex information-seeking tasks, producing evidence-grounded reports via multi-step reasoning, tool use, an

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents

DGX agent

arXiv:2605.18882v1 Announce Type: cross Abstract: LLM agents exhibit a consistent tendency to over-call, invoking tools even in situations where none is needed. On the When2Call benchmark, six models

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul Erdős in 1946. For nearly 80 …

DGX agent

Today, we share a breakthrough on the planar unit distance problem, a famous open question first posed by Paul Erdős in 1946. For nearly 80 years, mathematicians believed the best possible solutions l

model-releasesopenai--x
20 May 2026
Model Releases

Toto 2.0: Time Series Forecasting Enters the Scaling Era

DGX agent

arXiv:2605.20119v1 Announce Type: cross Abstract: We show that time series foundation models scale: a single training recipe produces reliable forecast-quality improvements from 4M to 2.5B parameters.

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Towards Camera-Robust 3D Localization: Equation-Anchored Tool-Use for MLLMs

DGX agent

arXiv:2605.19528v1 Announce Type: new Abstract: 3D localization in Multimodal Large Language Models (MLLMs), including 3D object detection and 3D visual grounding, is fundamentally limited by camera i

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

Towards Consistent Detection of Cognitive Distortions: LLM-Based Annotation and Dataset-Agnostic Evaluation

DGX agent

arXiv:2511.01482v2 Announce Type: replace Abstract: Text-based automated Cognitive Distortion detection is a challenging task due to its subjective nature, with low agreement scores observed even amon

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Towards Trust Calibration in Socially Interactive Agents: Investigating Gendered Multimodal Behaviors Generation with LLMs

DGX agent

arXiv:2605.19798v1 Announce Type: new Abstract: As Socially Interactive Agents (SIAs) become increasingly integrated into daily life, the ability to calibrate user trust to an agent's actual capabilit

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Training Neural Networks with Optimal Double-Bayesian Learning

DGX agent

arXiv:2605.20009v1 Announce Type: cross Abstract: Backpropagation with gradient descent is a common optimization strategy employed by most neural network architectures in machine learning. However, fi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Trajectory Planning and Control near the Limits: an Open Experimental Benchmark on the RoboRacer Platform

DGX agent

arXiv:2605.19881v1 Announce Type: new Abstract: We present a modular framework to benchmark new and existing methods for trajectory planning and control in high-acceleration maneuvers that push autono

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

TravExplorer: Cross-Floor Embodied Exploration via Traversability-Aware 3-D Planning

DGX agent

arXiv:2605.19958v1 Announce Type: new Abstract: Zero-shot Object Navigation (ZSON) has shown promise for open-vocabulary target search in unseen environments, yet most existing systems remain tied to

model-releasesarxiv-cs-ro
20 May 2026
Model Releases

.@trq212 is a builder's builder. After research at MIT, exiting his company, raising millions for another, and exploring research threads as…

DGX agent

.@trq212 is a builder's builder. After research at MIT, exiting his company, raising millions for another, and exploring research threads as an SPC member, he's now on the team building Claude Code. H

model-releasesthariq--x
20 May 2026
Model Releases

Trust or Abstain? A Self-Aware RAG Approach

DGX agent

arXiv:2605.18792v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) improves large language models (LLMs) by incorporating external evidence, but it also introduces knowledge confli

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

DGX agent

arXiv:2605.18859v1 Announce Type: cross Abstract: LLM routing matters most in long-horizon applications such as coding agents, deep research systems, and computer-use agents, where a single user reque

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Two research papers describe how Google's Co-Scientist and nonprofit FutureHouse's AI tools can succeed at drug-retargeting tasks by forming hypotheses (John Timmer/Ars Technica)

DGX agent

John Timmer / Ars Technica: Two research papers describe how Google's Co-Scientist and nonprofit FutureHouse's AI tools can succeed at drug-retargeting tasks by forming hypotheses — Both tools generat

model-releasestechmeme
20 May 2026
Model Releases

Unified Deployment-Aware Evaluation of Open Reasoning Language Models

DGX agent

arXiv:2604.07035v2 Announce Type: replace Abstract: Open reasoning language models are often compared under mixed sample sizes, partially standardized prompts, and accuracy-centered summaries, which m

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Unlocking the Potential of Continual Model Merging: An ODE Perspective

DGX agent

arXiv:2605.19409v1 Announce Type: cross Abstract: Continual Model Merging (CMM) enables rapid customization of foundation models across sequentially arriving tasks, offering a scalable alternative to

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Very interesting results from this NanoGPT-Bench eval. There is so much talk about self-improving agents. But can coding agents do real AI R…

DGX agent

Very interesting results from this NanoGPT-Bench eval. There is so much talk about self-improving agents. But can coding agents do real AI R&D? @IntologyAI reports that Codex, Claude Code, and Autores

model-releasesdair-ai--x
20 May 2026
Model Releases

ViroGym: Realistic Large-Scale Benchmarks for Evaluating Viral Proteins

DGX agent

arXiv:2603.06740v2 Announce Type: replace-cross Abstract: Protein language models (pLMs) have shown strong potential for zero-shot prediction of missense variant effects, yet systematic benchmarking o

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Vision Harnessing Agent for Open Ad-hoc Segmentation

DGX agent

arXiv:2605.19410v1 Announce Type: new Abstract: Segmentation has become easy when the concept is known, requiring retrieval of a learned visual grounding from text. It remains hard for open ad-hoc con

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗

DGX agent

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗 Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it t

model-releasesclem-delangue--x
20 May 2026
Model Releases

WARC-Bench: Web Archive Based Benchmark for GUI Subtask Executions

DGX agent

arXiv:2510.09872v2 Announce Type: replace-cross Abstract: Training web agents to navigate complex, real-world websites requires them to master extit{subtasks} - short-horizon interactions on multiple

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

We partnered with artists, designers, and builders to create new AI tools that solve real problems in their creative workflows. Here’s what’…

DGX agent

We partnered with artists, designers, and builders to create new AI tools that solve real problems in their creative workflows. Here’s what’s new: — Introducing Google Pics in @GoogleWorkspace: A bran

model-releasesgoogle-ai--x
20 May 2026
Model Releases

We're building Gemini for Science with and for the scientific community. In collaboration with 100+ institutions and a trusted tester commun…

DGX agent

We're building Gemini for Science with and for the scientific community. In collaboration with 100+ institutions and a trusted tester community that ranges from PhD students to Nobel laureates, we wan

model-releasesgoogle-ai--x
20 May 2026
Model Releases

We're excited to be an official shoutout at the Google I/O Developer Keynote 🔥 @llama_index is building the document infrastructure for AI …

DGX agent

We're excited to be an official shoutout at the Google I/O Developer Keynote 🔥 @llama_index is building the document infrastructure for AI agents, and we plan to integrate even more heavily with both

model-releasesjerry-liu--x
20 May 2026
Model Releases

We’re expanding our partnership with @SpaceX, and will be scaling up on GB200 capacity in Colossus 2 throughout June. Appreciate @elonmusk a…

DGX agent

We’re expanding our partnership with @SpaceX, and will be scaling up on GB200 capacity in Colossus 2 throughout June. Appreciate @elonmusk and the team helping us find good homes for the Claudes. In t

model-releaseselon-musk--x
20 May 2026
Model Releases

What Do Evolutionary Coding Agents Evolve?

DGX agent

arXiv:2605.20086v1 Announce Type: cross Abstract: Recent work pairs LLMs with evolutionary search to iteratively generate, modify, and select code using task-specific feedback. These systems have prod

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

What Makes a Representation Good for Single-Cell Perturbation Prediction?

DGX agent

arXiv:2605.19343v1 Announce Type: new Abstract: Single-cell perturbation modeling is fundamental for understanding and predicting cellular responses to genetic perturbations. However, existing approac

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

What Makes Synthetic Data Effective in Image Segmentation

DGX agent

arXiv:2605.19289v1 Announce Type: new Abstract: Driven by rapid advances in large-scale generative models, synthetic data has emerged as a promising solution for visual understanding. While modern dif

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

What we learned testing 7 models under the same agent harness

DGX agent

Model swaps look like configuration changes, but they behave more like product migrations. A new model may be cheaper, faster, easier to get capacity for, or stronger on public benchmarks.... The post

model-releasesarize-ai
20 May 2026
Model Releases

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection

DGX agent

arXiv:2601.22569v2 Announce Type: replace-cross Abstract: Large language model (LLM) based agents are increasingly used to automate financial transactions, yet their reliance on contextual reasoning e

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

World-Ego Modeling for Long-Horizon Evolution in Hybrid Embodied Tasks

DGX agent

arXiv:2605.19957v1 Announce Type: cross Abstract: World models are widely explored in embodied intelligence, yet they typically predict distinct evolutions of the world and the ego within a single str

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Worldlier: native support for 48 world languages and improved efficiency in non-European languages.

DGX agent

Worldlier is a Cohere language model that natively supports 48 world languages with improved efficiency, particularly for non-European languages. This represents an expansion of language coverage beyo

model-releasescohere--x
20 May 2026
Model Releases

WoundFormer: Multi-Scale Spatial Feature Fusion for Multi-Class Wound Tissue Segmentation

DGX agent

arXiv:2605.19868v1 Announce Type: new Abstract: Chronic wounds such as diabetic foot ulcers and pressure injuries require accurate tissue-level assessment to guide treatment planning and monitor heali

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

XNote: Benchmarking Automated Community Notes Generation for Image-based Contextual Deception

DGX agent

arXiv:2603.22453v2 Announce Type: replace Abstract: Community Notes have emerged as an effective crowd-sourced mechanism for combating online deception on social media platforms. However, its reliance

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

You can now remix other people’s YouTube Shorts with AI

DGX agent

Google announced a new YouTube Shorts Remix feature that lets users restyle clips or even insert themselves into other people's videos using Gemini Omni. Now, at the bottom of a YouTube Short, when yo

model-releasesthe-verge-ai
20 May 2026
Model Releases

ZeroSearch: Incentivize the Search Capability of LLMs without Searching

DGX agent

arXiv:2505.04588v3 Announce Type: replace Abstract: Effective information searching is essential for enhancing the reasoning and generation capabilities of large language models (LLMs). Recent researc

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

ZeroUnlearn: Few-Shot Knowledge Unlearning in Large Language Models

DGX agent

arXiv:2605.18879v1 Announce Type: cross Abstract: Large language models inevitably retain sensitive information, defined as inputs that may induce harmful generations, due to training on massive web c

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

1GC-7RC: One Graphic Card -- Seven Research Challenges! How Good Are AI Agents at Doing Your Job?

DGX agent

arXiv:2605.17046v1 Announce Type: cross Abstract: Autonomous AI coding agents are becoming a core tool for ML practitioners in industry and research alike. Despite this growing adoption, no standardiz

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

3DPhysVideo: Consistency-Guided Flow SDE for Video Generation via 3D Scene Reconstruction and Physical Simulation

DGX agent

arXiv:2605.16795v1 Announce Type: cross Abstract: Video generative models have made remarkable progress, yet they often yield visual artifacts that violate grounding in physical dynamics. Recent works

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…290291292293294…472
Next →