AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,378 results
9 Apr 2026

Silicon Valley is quietly running on Chinese open source AI models. Here are the receipts: → Cursor confirmed last month that Composer 2 is …

Model ReleasesDGX agent

Silicon Valley is quietly running on Chinese open source AI models. Here are the receipts: → Cursor confirmed last month that Composer 2 is built on Moonshot's Kimi K2.5 → Cognition's SWE-1.6 model is

Simplicity is a very strong signal of model quality. Galileo's heliocentric model was directionally correct, but its predictive power was qu…

ResearchDGX agent

Simplicity is a very strong signal of model quality. Galileo's heliocentric model was directionally correct, but its predictive power was quite bad compared to the much older Ptolemaic model, because

The US frontier labs have all walked away from open weights. They continue to occasionally release excellent open models (Gemma 4, etc), but…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

The US frontier labs have all walked away from open weights. They continue to occasionally release excellent open models (Gemma 4, etc), but they are smaller models that are not competitive with their

8 Apr 2026

New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped analysis! 8/8 models found the flags…

ResearchDGX agent

New post: We tested the Mythos showcase vulnerabilities with open models. They recovered similar scoped analysis! 8/8 models found the flagship FreeBSD zero-day, including a 3B model. Rankings reshuff

7 Apr 2026

With SWE-1.6 we've made significant progress on 'intelligence per token'. We post-trained the model from scratch (same pre-trained model) wi…

AgentsDGX agent

With SWE-1.6 we've made significant progress on 'intelligence per token'. We post-trained the model from scratch (same pre-trained model) with a similar recipe as SWE-1.6 Preview. Our latest algorithm

12 Aug 2026

Today, we’re adding another member to our model family. Meet North Micro Vision. Our smallest vision-language model yet, ideal for sophistic…

Model ReleasesDGX agent

Today, we’re adding another member to our model family. Meet North Micro Vision. Our smallest vision-language model yet, ideal for sophisticated document understanding. Available open-source under an

11 Aug 2026

It's been exciting for Ollama to partner with @JensenHuang and the @NVIDIAAI team on launching open models. Open models have no boundaries, …

Model ReleasesDGX agent

It's been exciting for Ollama to partner with @JensenHuang and the @NVIDIAAI team on launching open models. Open models have no boundaries, and let's continue to work together to make this ecosystem b

Forgetting-Resistant and Lesion-Aware Source-Free Domain Adaptive Fundus Image Analysis with Vision-Language Model

Model ReleasesDGX agent

arXiv:2602.19471v2 Announce Type: replace Abstract: Source-free domain adaptation (SFDA) aims to adapt a model trained in the source domain to perform well in the target domain, with only unlabeled ta

10 Aug 2026

Claude Haiku is my current least favorite model - it hallucinates wildly, and is out-performed now by other similarly priced models like GPT…

Model ReleasesDGX agent

Claude Haiku is my current least favorite model - it hallucinates wildly, and is out-performed now by other similarly priced models like GPT-5.6-Luna Even worse: it seems to still be used by the Claud

Meta released Muse Glimmer 30B: multimodal model for your Claw/Pi setups 🔥 we tested and fine-tuned the model for you, and shipped day-0 su…

Model ReleasesDGX agent

Meta released Muse Glimmer 30B: multimodal model for your Claw/Pi setups 🔥 we tested and fine-tuned the model for you, and shipped day-0 support in transformers and llama.cpp, including DFlash for 2-4

31 Jul 2026

Experience sharing: How do you use your local models and for what kind of tasks?

Model ReleasesDGX agent

Here is my experience, which I would like to share with you and I also would like to hear your thoughts and valuable tips&tricks. Hardware: Mac Mini M4 (32GB Unified Memory) Model Server: Ollama Orche

Has anyone actually benchmarked where the 'big-model orchestrator + local-model worker' split breaks down?

Model ReleasesDGX agent

I keep seeing the 'use a big model via API as the architect, run local small/mid models as workers' pattern recommended for people with modest local hardware. I've been running it myself (orchestrator

Selecting Open-Weight Language Models for Zero-Shot Intent Classification: A Systematic Evaluation of 41 Models

Model ReleasesDGX agent

arXiv:2607.27421v1 Announce Type: new Abstract: Intent classification is a core component of task-oriented dialogue systems, yet practitioners have limited systematic guidance for selecting deployable

Uncensored Multi-Model Releases, LongCat-Flash-Lite with MTPs, Jamba2-Mini, Qwen3.5-9B-Nikusui-v1 with MTPs and Qwen3.5-27B-Nikusui-v1 with MTPs, Available in Safetensors and GGUF Formats!

Model ReleasesDGX agent

Been working hard for the past month to bring to the community some interesting curios, so for starters we have LongCat-Flash-Lite Uncensored Heretic with MTPs which has never before been uncensored,

28 Jul 2026

Modeling Memory-Dependent Reliability of LLMs: A Hidden Markov Model

Model ReleasesDGX agent

arXiv:2607.22951v1 Announce Type: cross Abstract: Reliability assessment of large language models (LLMs) seeks to estimate the probability that a model produces correct responses under a specified ope

Why Anthropic's battle is meant to poison the wells of open weight models, in 3 steps.

SafetyDGX agent

It doesn't solve any problems. Just a few paragraphs above, he says he fears that authoritarian states (he names China, and possibly others) can use their models to do evil stuff. And surely enough, m

Rethinking Expert Training for Model Merging with Prompt Learning

Model ReleasesDGX agent

arXiv:2607.24465v1 Announce Type: new Abstract: Model merging aims to combine multiple domain-specialized experts trained from a shared foundation model into a single multi-task model. Existing approa

15 Jul 2026

Thinky with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), new best American model…

Model ReleasesDGX agent

Thinky with a ~1T param, 41B active, apache-2 model Benchmarks are a clear step up from Nemotron Ultra (55B active), new best American model, and omni input. A bit behind GLM 5.2 on agentic benchies,

8 Jul 2026

„Language models model language“ is going to be my quote of the week … maybe I print a T shirt for the next conference.

SafetyDGX agent

„Language models model language“ is going to be my quote of the week … maybe I print a T shirt for the next conference. super interesting - and a reminder that language models model language and not i

1 Jul 2026

Fable, the indomitable model from Anthropic, is coming back. These last two weeks have been a great reminder that: models can be removed at …

Model ReleasesDGX agent

Fable, the indomitable model from Anthropic, is coming back. These last two weeks have been a great reminder that: models can be removed at any time, you should have fallback plans (including open sou

Learning by Surprise: Adaptive Mitigation of Model Collapse in Large Language Models

ResearchDGX agent

arXiv:2410.12341v4 Announce Type: replace-cross Abstract: As AI-generated content increasingly populates the web, generative AI models are at growing risk of being trained on their own outputs, a proc

TDGT: A Tabular Data Generation Toolkit supporting adaptive GPU-accelerated Bayesian mixture models, diffusion-based models, and latent-space generative modeling

SafetyDGX agent

arXiv:2606.31268v1 Announce Type: cross Abstract: The growing demand for privacy-preserving data sharing has positioned synthetic data generation as a critical component of responsible AI workflows. D

30 Jun 2026

LWDrive: Layer-Wise World-Model-Guided Vision-Language Model Planning for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.29879v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) provide powerful semantic understanding and commonsense reasoning for End-to-End Autonomous Driving (E2E-AD) planning. H

Stochastic and Non-local Closure Modeling for Nonlinear Dynamical Systems via Latent Score-based Generative Models

Local AiDGX agent

arXiv:2506.20771v2 Announce Type: replace Abstract: We propose a latent score-based generative AI framework for learning stochastic, non-local closure models and constitutive laws in nonlinear dynamic

26 Jun 2026

Joint Reward Modeling: Internalizing Chain-of-Thought for Efficient Visual Reward Models

Local AiDGX agent

arXiv:2602.07533v2 Announce Type: replace Abstract: Reward models are critical for reinforcement learning from human feedback, as they determine the alignment quality and reliability of generative mod

23 Jun 2026

Tesla Model 3 and Model Y have the highest percentage of American-made content!

IndustryDGX agent

Tesla Model 3 and Model Y have the highest percentage of American-made content! The @Tesla Model 3 and Model Y have just been named the #1 and #2 Most American-Made vehicles for 2026 by http://Cars.co

7 Jun 2026

Super-powerful AI models will launch in the coming weeks. We are looking at a potential step change in model capabilities. The biggest mista…

TutorialsDGX agent

Super-powerful AI models will launch in the coming weeks. We are looking at a potential step change in model capabilities. The biggest mistake right now is to lock into one vendor. I say this not only

5 Jun 2026

Do Models Share Safety Representations? Cross-Model Steering for Safe Visual Generation

SafetyDGX agent

arXiv:2606.05290v1 Announce Type: new Abstract: Recent progress in generative modeling has made safety control a central challenge, yet existing approaches remain largely model-specific, requiring ret

4 Jun 2026

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running cod…

Model ReleasesDGX agent

NEW: NVIDIA ships 550B MoE open model for long-running agents. Very exciting times to see more open models to support local long-running coding agents. Today we're shipping Nemotron 3 Ultra. A 550B Mo

The Variance Brain Foundation Models Forgot: Third-Order Statistics Predict Cognition Where Billion-Parameter Models Fail

Model ReleasesDGX agent

arXiv:2606.04010v1 Announce Type: cross Abstract: Brain foundation models (BFMs) are self-supervised Transformers pretrained on fMRI data. We posit that these models should capture each subject's cogn

Unifying Model-Free Efficiency and Model-Based Representations via Latent Dynamics

Model ReleasesDGX agent

arXiv:2602.12643v2 Announce Type: replace-cross Abstract: We present Unified Latent Dynamics (ULD), a novel reinforcement learning algorithm that unifies the efficiency of model-free methods with the

3 Jun 2026

Using open models and inference clouds (which serve open models) is a leading indicator of what is to come. The advantage of open weights is…

Model ReleasesDGX agent

Using open models and inference clouds (which serve open models) is a leading indicator of what is to come. The advantage of open weights is that you can train, serve, and continually improve your own

21 May 2026

i feel like there's a general misunderstanding about open source models. most people use a frontier model, switch the api request to open so…

AgentsDGX agent

i feel like there's a general misunderstanding about open source models. most people use a frontier model, switch the api request to open source model, see poor performance, and then churn off. this w

12 May 2026

Are vision-language models ready to zero-shot replace supervised classification models in agriculture?

Model ReleasesDGX agent

arXiv:2512.15977v3 Announce Type: replace Abstract: Vision-language models (VLMs) are increasingly proposed as general-purpose solutions for visual recognition tasks, yet their reliability for agricul

11 May 2026

ModelLens: Finding the Best for Your Task from Myriads of Models

Model ReleasesDGX agent

arXiv:2605.07075v1 Announce Type: new Abstract: The open-source model ecosystem now contains hundreds of thousands of pretrained models, yet picking the best model for a new dataset is increasingly in

5 May 2026

Rethinking the Need for Source Models: Source-Free Domain Adaptation from Scratch Guided by a Vision-Language Model

TutorialsDGX agent

arXiv:2605.02604v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) adapts source models to target domains without accessing source data, addressing privacy and transmission issues. H

The Partial Testimony of Logs: Evaluation of Language Model Generation under Confounded Model Choice

SafetyDGX agent

arXiv:2605.01311v1 Announce Type: new Abstract: Offline evaluation of language models from usage logs is biased when model choice is confounded: the same user-side factors that influence which model i

4 May 2026

Not every step in an agent workflow needs the same model. Fleet now lets you customize which model each sub-agent uses, so you can route sim…

AgentsDGX agent

Not every step in an agent workflow needs the same model. Fleet now lets you customize which model each sub-agent uses, so you can route simple tasks to fast/cheap models and keep stronger models for

1 May 2026

Can Large Language Models Implement Agent-Based Models? An ODD-based Replication Study

Model ReleasesDGX agent

arXiv:2602.10140v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can now synthesize non-trivial executable code from textual descriptions, raising an important question: can LLMs

28 Apr 2026

I’ve been saying this for over a year now: Frontier models are fantastic, but the real future is frontier-level models (in every way) runnin…

Model ReleasesDGX agent

I’ve been saying this for over a year now: Frontier models are fantastic, but the real future is frontier-level models (in every way) running locally on your own hardware. I think this is 18-24 months

SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems

Model ReleasesDGX agent

arXiv:2604.23905v1 Announce Type: cross Abstract: Threat modeling for cyber-physical systems (CPS) remains a largely manual exercise. This project presents SMSI (System Model Security Inference), a hy

27 Apr 2026

Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!

Model ReleasesDGX agent

arXiv:2507.03014v2 Announce Type: replace-cross Abstract: Large language models (LLMs) face significant copyright and intellectual property challenges as the cost of training increases and model reuse

Mapping a smarter future with BigQuery and Google Earth AI models and datasets

Model ReleasesDGX agent

Last year we introduced new geospatial analytics capabilities integrated for BigQuery. Building on this, we announced an expanded suite of tools at Google Cloud Next ‘26, designed to help your busines

Multi-output Extreme Spatial Model for Complex Aircraft Production Systems

Model ReleasesDGX agent

arXiv:2604.22548v1 Announce Type: cross Abstract: Problem definition: Data-driven models in machine learning have enabled efficient management of production systems. However, a majority of machine lea

24 Apr 2026

And now a new DeepSeek model, and appears to be fully open weights. Good benchmarks, but with open models, that isn't always as meaningful. …

Model ReleasesDGX agent

DeepSeek released a new open-weights model with strong benchmark performance, though Mollick notes that benchmark results may not fully capture the capabilities of open models compared to closed syste

DeepSeek V4 Pro has 1.6T total parameters, its largest model by the metric, and V4 Flash has 284B parameters; both models have a context window of 1M tokens (Vincent Chow/South China Morning Post)

Model ReleasesDGX agent

Vincent Chow / South China Morning Post: DeepSeek V4 Pro has 1.6T total parameters, its largest model by the metric, and V4 Flash has 284B parameters; both models have a context window of 1M tokens —

Learning Physics from Pretrained Video Models: A Multimodal Continuous and Sequential World Interaction Models for Robotic Manipulation

SafetyDGX agent

arXiv:2603.00110v2 Announce Type: replace Abstract: The scarcity of large-scale robotic data has motivated the repurposing of foundation models from other modalities for policy learning. In this work,

23 Apr 2026

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training

SafetyDGX agent

arXiv:2604.20012v1 Announce Type: cross Abstract: Vision-Language-Action Models (VLAs) inherit their visual and linguistic capabilities from Vision-Language Models (VLMs), yet most VLAs are built from

20 Apr 2026

Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k…

Model ReleasesDGX agent

Kimi K2.6 raises the bar for open-source models. 🦙 available on Ollama's cloud! Try it with OpenClaw: ollama launch openclaw --model kimi-k2.6:cloud Try it with Hermes Agent: ollama launch hermes --mo

16 Apr 2026

Robust Reward Modeling for Large Language Models via Causal Decomposition

Model ReleasesDGX agent

arXiv:2604.13833v1 Announce Type: new Abstract: Reward models are central to aligning large language models, yet they often overfit to spurious cues such as response length and overly agreeable tone.

Selecting Feature Interactions for Generalized Additive Models by Distilling Foundation Models

TutorialsDGX agent

arXiv:2604.13332v1 Announce Type: new Abstract: Identifying meaningful feature interactions is a central challenge in building accurate and interpretable models for tabular data. Generalized additive

15 Apr 2026

Banger paper from NVIDIA. Agentic reasoning needs models that are not just capable, but efficient at long-context inference. The agent model…

Model ReleasesDGX agent

Banger paper from NVIDIA. Agentic reasoning needs models that are not just capable, but efficient at long-context inference. The agent model layer is moving toward open, long-context, high-throughput

14 Apr 2026

Deliberative Alignment is Deep, but Uncertainty Remains: Inference time safety improvement in reasoning via attribution of unsafe behavior to base model

SafetyDGX agent

arXiv:2604.09665v1 Announce Type: cross Abstract: While the wide adoption of refusal training in large language models (LLMs) has showcased improvements in model safety, recent works have highlighted

Learning World Models for Interactive Video Generation

Model ReleasesDGX agent

arXiv:2505.21996v3 Announce Type: replace-cross Abstract: Foundational world models must be both interactive and preserve spatiotemporal coherence for effective future planning with action choices. Ho

Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models

ResearchDGX agent

arXiv:2604.02340v2 Announce Type: replace-cross Abstract: Recent advances in masked diffusion language models (MDLMs) narrow the quality gap to autoregressive LMs, but their sampling remains expensive

13 Apr 2026

Sentiment Classification of Gaza War Headlines: A Comparative Analysis of Large Language Models and Arabic Fine-Tuned BERT Models

Model ReleasesDGX agent

arXiv:2604.08566v1 Announce Type: new Abstract: This study examines how different artificial intelligence architectures interpret sentiment in conflict-related media discourse, using the 2023 Gaza War

7 Aug 2026

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in thi…

Model ReleasesDGX agent

Codex is overoptimised for large models: it ranks 2nd out of 10 for GLM 5.2 but drops to 9th place for Gemma-4! Almost all the effort in this field goes into tuning the weights. We wanted to know how

24 Jul 2026

A new model launch is not a product update ‼ It's one of three things, and you don't know which until you test it. Sometimes it's nothing: t…

Model ReleasesDGX agent

A new model launch is not a product update ‼ It's one of three things, and you don't know which until you test it. Sometimes it's nothing: the model improved inside the same distribution, your harness

29 Jun 2026

Supercharging the agentic era with Spanner’s multi-model architecture

Model ReleasesDGX agent

In the agentic era, the role of the database has fundamentally changed. It is no longer a passive repository; it’s a critical context engine designed to ground generative AI apps, models and power aut

2 Jun 2026

Sharpness-Aware Hybrid Model Learning for Architecture-Agnostic Parameter Estimation

Model ReleasesDGX agent

arXiv:2602.06837v2 Announce Type: replace Abstract: Hybrid modeling, the combination of machine learning models and scientific mathematical models, enables flexible and robust data-driven prediction w

← Previous
123456…990
Next →