AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
Human
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61+ results
20 Jun 2026

🚨🚨🚨A research project idea! How to measure world models? Everyone's talking about world models these days. World model here, world model …

AgentsDGX agent

🚨🚨🚨A research project idea! How to measure world models? Everyone's talking about world models these days. World model here, world model there. We can argue about what 'world model' actually means, an

22 Apr 2026

Qwen 3.6 27B model is available on Ollama! Use it with all the integrations in Ollama or chat with the model. Chat with the model: ollama ru…

Model ReleasesDGX agent

Qwen 3.6 27B model is available on Ollama! Use it with all the integrations in Ollama or chat with the model. Chat with the model: ollama run qwen3.6:27b OpenClaw: ollama launch openclaw --model qwen3

🦤 LeWorldModel: Learning Physics from Pixels — Stable World Models with Just Two Losses World models: 1️⃣ DINO-WM: pretrained ViT encoder (…

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

🦤 LeWorldModel: Learning Physics from Pixels — Stable World Models with Just Two Losses World models: 1️⃣ DINO-WM: pretrained ViT encoder (from ImageNet) → features → predictor. But encoder is frozen,

Qwen3.6 and gemma 4 just shown that we were still far from capability limit on small models, we still are. Huge models are good, but is ther…

Model ReleasesDGX agent

Qwen3.6 and gemma 4 just shown that we were still far from capability limit on small models, we still are. Huge models are good, but is there really a point in going bigger. Just realize something. On

i haven't seen a model that just works across agent harnesses. seems like it should exist. great opportunity for open-weight models. any tho…

AgentsDGX agent

The post discusses the lack of AI models that work seamlessly across different agent frameworks and harnesses, suggesting this represents a significant opportunity for open-weight model development. T

19 Apr 2026

Who is running local models on GPUs on OpenClaw? I have started benchmarking different models this week. I am working on improving model sel…

Model ReleasesDGX agent

Who is running local models on GPUs on OpenClaw? I have started benchmarking different models this week. I am working on improving model selection and switching UX on OpenClaw, i.e. I run /model vllm/

11 Apr 2026

Hermes Agent users are remarkably model-curious. Of the top 10 models used in Hermes Agent, 8 different model companies are represented, and…

AgentsDGX agent

Hermes Agent users are remarkably model-curious. Of the top 10 models used in Hermes Agent, 8 different model companies are represented, and 6 of the top 7 are open source. Without lock-in, people try

29 Apr 2026

Model-Harness-Task fit is very real. The world's best agents carefully tailor the harness around the model to take advantage of each model's…

Model ReleasesDGX agent

Model-Harness-Task fit is very real. The world's best agents carefully tailor the harness around the model to take advantage of each model's unique intelligence & capabilities 🧠 Today we released Harn

Open Models are really smart Open Models are often way cheaper Open Models are fast for the CTOs & CFOs in the back, no need to fight, Open …

AgentsDGX agent

Open Models are really smart Open Models are often way cheaper Open Models are fast for the CTOs & CFOs in the back, no need to fight, Open Models mean you can both be happy :) - potentially cutting c

27 Jul 2026

Releasing the model weights and technical report of Kimi K3. Kimi K3 is our most capable model: a 2.8T MoE model with native visual understa…

Model ReleasesDGX agent

Releasing the model weights and technical report of Kimi K3. Kimi K3 is our most capable model: a 2.8T MoE model with native visual understanding and a 1M-token context window. New model architecture:

27 May 2026

Model-Harness-Task fit! it’s clear that RL post-training produces a model-harness fit via tool shapes and prompting as models are trained wi…

Model ReleasesDGX agent

Model-Harness-Task fit! it’s clear that RL post-training produces a model-harness fit via tool shapes and prompting as models are trained with the harness in the loop. Mentioned this in a previous Lan

24 Apr 2026

New Frontier LLM Orchestrator Model @SakanaAILabs The fugu model series provides a novel paradigm for test-time scaling. The models are trai…

AgentsDGX agent

New Frontier LLM Orchestrator Model @SakanaAILabs The fugu model series provides a novel paradigm for test-time scaling. The models are trained to call LLMs and infer not only which individual model i

5 Aug 2026

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all…

Model ReleasesDGX agent

Document OCR is not Getting Commoditized (by Frontier Models) The most common question I get is whether frontier models are going to eat all document processing solutions - just screenshot the page an

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available …

Model ReleasesDGX agent

DeepSeek V4 Flash is Ollama's fastest growing model ever in token usage, and the most popular model on OpenRouter this week. It’s available in Pi across a range of providers. If you’ve never tried an

14 Apr 2026

The idea of sovereign frontier models is only viable as long as the main Chinese model makers keep shipping open weights models that anyone …

ApplicationsDGX agent

The idea of sovereign frontier models is only viable as long as the main Chinese model makers keep shipping open weights models that anyone can build on. I am not sure how long that will continues but

Harness Engineering Derived from what Models can’t do alone: It feels like a good time to step back and reshare some basic mental models for…

AgentsDGX agent

Harness Engineering Derived from what Models can’t do alone: It feels like a good time to step back and reshare some basic mental models for why harnesses exist in the first place - working backwards

9 Apr 2026

There are more competitive small model makers, but there is still a very big gap between what small models can do and what large models can …

ApplicationsDGX agent

There are more competitive small model makers, but there is still a very big gap between what small models can do and what large models can accomplish (even if the small model benchmarks say otherwise

So we now have a pretty good picture of the state of the frontier AI model makers. US closed source models continue to lead. Google, OpenAI,…

Model ReleasesDGX agent

So we now have a pretty good picture of the state of the frontier AI model makers. US closed source models continue to lead. Google, OpenAI, and Anthropic stand well ahead of the pack, and may have si

Routing every task to your largest model burns tokens, adds latency, and inflates costs. @AI21Labs' Maestro Orchestration Meta Model (OMM) i…

Model ReleasesDGX agent

Routing every task to your largest model burns tokens, adds latency, and inflates costs. @AI21Labs' Maestro Orchestration Meta Model (OMM) is the layer above your stack that dynamically selects the ri

9 Jul 2026

The model has an advisor tool that natively escalates to a stronger model when needed. This model is hosted in the U.S. by Perplexity on Nvi…

HardwareDGX agent

The model has an advisor tool that natively escalates to a stronger model when needed. This model is hosted in the U.S. by Perplexity on Nvidia B200 GPUs. We will improve the model in research preview

9 Aug 2026

The best 'raw' frontier model for document parsing is gemini 3 flash, but the issue is that since then the flash models have gotten 3x more …

Model ReleasesDGX agent

The best 'raw' frontier model for document parsing is gemini 3 flash, but the issue is that since then the flash models have gotten 3x more expensive while flatlining on visual recognition across comp

1 Aug 2026

Banning open source model, will 'remove capabilities for the defenders' When OpenAI's models breached Huggingface, a Chinese open model is w…

HardwareDGX agent

Banning open source model, will 'remove capabilities for the defenders' When OpenAI's models breached Huggingface, a Chinese open model is what cleaned up the mess. Anthropic's Fable 5 refused, so Hug

Your design, your model. Create with your pick of the world's leading models, including Claude, GPT-5, Gemini, Kimi, and GLM. Compare output…

Model ReleasesDGX agent

Your design, your model. Create with your pick of the world's leading models, including Claude, GPT-5, Gemini, Kimi, and GLM. Compare outputs across model families and keep the result that nails it. O

3 Jul 2026

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on …

Model ReleasesDGX agent

Don't train the model, evolve the harness. I read a brilliant blog post from Hugging Face where they took a frozen open model scoring 0% on a hard legal agent benchmark, left its weights alone, and le

10 Apr 2026

Sub-agent Model Selection — Different Tasks, Different Models Your main agent runs Qwen3.6-Plus for quality. But not every subtask needs a f…

AgentsDGX agent

Sub-agent Model Selection — Different Tasks, Different Models Your main agent runs Qwen3.6-Plus for quality. But not every subtask needs a flagship model. Now sub-agents can use a different model. Cre

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, …

Model ReleasesDGX agent

llama.cpp now supports various small OCR models that can run on low-end devices. These models are small enough to run on GPU with 4GB VRAM, and some of them can even run on CPU with decent performance

7 Apr 2026

We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview model on SWE-Bench Pro while dramatical…

AgentsDGX agent

We’re releasing SWE-1.6, our best model in both intelligence & model UX. SWE-1.6 matches our Preview model on SWE-Bench Pro while dramatically improving on various behavioral axes. It’s available toda

GLM-5.1 by @Zai_org just launched in the Text Arena, and is now the #1 open model. It outperforms the next best open model, its predecessor,…

Model ReleasesDGX agent

GLM-5.1 by @Zai_org just launched in the Text Arena, and is now the #1 open model. It outperforms the next best open model, its predecessor, GLM-5, by +11 points and +15 over Kimi K2.5 Thinking. It sh

2 Jun 2026

Open-weight models have overtaken closed models on OpenRouter. 69.1% of token volume now goes to open-weight models. 30.9% to closed. Compet…

IndustryDGX agent

Open-weight models have overtaken closed models on OpenRouter. 69.1% of token volume now goes to open-weight models. 30.9% to closed. Competition is a discovery procedure — and developers are discover

6 Jul 2026

Video models are the path to world models. The ability to build frontier world models will become one of the most important sources of compe…

IndustryDGX agent

Video models are the path to world models. The ability to build frontier world models will become one of the most important sources of competitive advantage for every country and every continent. JUST

4 Jun 2026

locked n loaded🔒 Made in @comfyUI, @AdobeAE + @Photoshop, @suno Vid models: @Alibaba_Wan 2.2, @ltx_model 2.3 @thesystms FLW Img models: @op…

Local AiDGX agent

locked n loaded🔒 Made in @comfyUI, @AdobeAE + @Photoshop, @suno Vid models: @Alibaba_Wan 2.2, @ltx_model 2.3 @thesystms FLW Img models: @openai GPT 2.0, @googleai Nano Banana 2, @krea 2, @grok GFX: @t

24 Jul 2026

Open models matter. Ollama works hard with the model creators, hardware partners, and most importantly developers building software leveragi…

Local AiDGX agent

Open models matter. Ollama works hard with the model creators, hardware partners, and most importantly developers building software leveraging various open models for their own use cases. For my first

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models.

SafetyDGX agent

Welcome @jensenhuang 💚 The future of AI leadership needs both frontier closed models and frontier open models. For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will

27 Apr 2026

Model page: https://ollama.com/library/deepseek-v4-pro Try it with Codex: ollama launch codex --model deepseek-v4-pro:cloud Try it with Open…

Model ReleasesDGX agent

Model page: https://ollama.com/library/deepseek-v4-pro Try it with Codex: ollama launch codex --model deepseek-v4-pro:cloud Try it with OpenClaw: ollama launch openclaw --model deepseek-v4-pro:cloud T

Hermes will now by default use native vision if the main agent model supports it, and you didn't set a different vision auxiliary model! Jus…

AgentsDGX agent

Hermes will now by default use native vision if the main agent model supports it, and you didn't set a different vision auxiliary model! Just `hermes update` and it will take affect immediately. You c

3 Aug 2026

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... bu…

Model ReleasesDGX agent

You probably used @FireworksAI_HQ this week. The cool part is that you just did not know you did. 🎆 Open-source models are free, sure... but the hard part is tailoring them so they perform best on you

12 Jul 2026

playing with local AI models and doing some research and came across a fun extrapolation: 'consumer grade graphics cards will be running Fab…

Model ReleasesDGX agent

playing with local AI models and doing some research and came across a fun extrapolation: 'consumer grade graphics cards will be running Fable-equivalent models by 2029' the argument: 1) been said tha

22 Jun 2026

Another new idea to push the state of AI architectures forward. Sakana released a model that effectively uses a mixture of models to get wor…

Model ReleasesDGX agent

Another new idea to push the state of AI architectures forward. Sakana released a model that effectively uses a mixture of models to get work done. You get a single API but then the work gets farmed o

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference betwe…

SafetyDGX agent

have been thinking a bunch about model routing and related things current thoughts here, would love feedback: 1/ there is a difference between 'model routing' and 'model council' 'model routing' = rou

1 Jun 2026

Big day for American open models... Nemotron 3 Ultra is now the strongest US open-weight model tested, while apparently serving 300+ tok/s …

Model ReleasesDGX agent

Big day for American open models... Nemotron 3 Ultra is now the strongest US open-weight model tested, while apparently serving 300+ tok/s 🤯 Comparable large DeepSeek/Kimi models are usually 50-100 to

.@MiniMax_AI M3 model is available on Ollama's Cloud! In partnership with MiniMax, the M3 model on Ollama's Cloud is US-based with zero data…

Model ReleasesDGX agent

.@MiniMax_AI M3 model is available on Ollama's Cloud! In partnership with MiniMax, the M3 model on Ollama's Cloud is US-based with zero data retention. Try M3 on coding and agentic tasks: Claude Code:

27 Jun 2026

Run Ornith with Ollama: ollama run ornith For coding, use it with Claude or Pi: ollama launch claude --model ornith ollama launch pi --model…

Model ReleasesDGX agent

Run Ornith with Ollama: ollama run ornith For coding, use it with Claude or Pi: ollama launch claude --model ornith ollama launch pi --model ornith For the more capable 35B model, use: ollama launch c

Thanks for running our open-source work on current frontier models “The results are: the most capable models today (GPT-5.5 Pro) did outperf…

Model ReleasesDGX agent

Thanks for running our open-source work on current frontier models “The results are: the most capable models today (GPT-5.5 Pro) did outperform the best models from before (79/100 vs 69/100), but did

7 Aug 2026

After evaluating one of our upcoming models, Astra, we're treating it as our first 'critical' model for cybersecurity under our Preparedness…

Model ReleasesDGX agent

After evaluating one of our upcoming models, Astra, we're treating it as our first 'critical' model for cybersecurity under our Preparedness Framework. This is a scenario we've planned for, and we're

4 Aug 2026

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model …

Model ReleasesDGX agent

DeepSeek-V4-Flash-0731 is Ollama's fastest growing model ever in token usage. We are scaling capacity in US & Europe. On Ollama, this model runs with high performance (100tps+) and zero data retention

10 Jul 2026

🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a …

Model ReleasesDGX agent

🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a NemoClaw DeepAgents blueprint. This pairs Deep Agents (our op

26 Jun 2026

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from ex…

Model ReleasesDGX agent

UBS says 60% of companies now watching AI budgets are moving to cheaper models and open-source Chinese models The pressure is coming from extreme bills, including users spending up to $35K/month, team

Introducing a limited preview of GPT-5.6 Sol, our next generation frontier model, as well as GPT-5.6 Terra, a balanced model for efficient, …

Model ReleasesDGX agent

Introducing a limited preview of GPT-5.6 Sol, our next generation frontier model, as well as GPT-5.6 Terra, a balanced model for efficient, everyday work, and GPT-5.6 Luna, a fast and affordable model

25 Jun 2026

Video models are hard. Real-time video models are even harder. Loved this breakdown by @itunpredictable on how we productionized Runway Char…

Model ReleasesDGX agent

Video models are hard. Real-time video models are even harder. Loved this breakdown by @itunpredictable on how we productionized Runway Characters, our real-time interactive avatar model. We had to fi

21 Jun 2026

Very impressive from GLM-5.2. Frontier open-weight model indeed. Now, can we get a Gemini model in the top 3 soon?

Model ReleasesDGX agent

Very impressive from GLM-5.2. Frontier open-weight model indeed. Now, can we get a Gemini model in the top 3 soon? GLM 5.2 is now on DeepSWE as the top open-source model on our leaderboard. With a pas

9 Jun 2026

NEW: Anthropic introduces Claude Fable 5, a Mythos-class model for general use. Beginning of a new class of frontier models.

Model ReleasesDGX agent

NEW: Anthropic introduces Claude Fable 5, a Mythos-class model for general use. Beginning of a new class of frontier models. Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for g

26 May 2026

@mteamisloading the models from 6 months ago kinda feel the same like the recently released models. currently not holding my breath for more…

Model ReleasesDGX agent

Jeremy Howard comments that large language models released 6 months ago feel comparable in capability to recently released models, suggesting that the pace of improvement in model development may be s

6 May 2026

we're continuing to see clear examples where a model's harness is a major determinant of overall performance. with the same model, running o…

Model ReleasesDGX agent

we're continuing to see clear examples where a model's harness is a major determinant of overall performance. with the same model, running on same task, it's easy to observe very different scores depe

30 Apr 2026

Mythos seems to be a very capable model based on available information, but it is not a cybersecurity model - it is an advanced general purp…

ApplicationsDGX agent

Mythos seems to be a very capable model based on available information, but it is not a cybersecurity model - it is an advanced general purpose model that happens to be good at cyber because it is goo

28 Apr 2026

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page

Model ReleasesDGX agent

First open-weight model from @poolsideai! Apache license, and available on Ollama to try. 👇👇👇 model page Today we’re releasing Laguna XS.2, Poolside’s first open-weight model. It’s a 33B total / 3B ac

17 Apr 2026

We need a new document that AI labs should release with each new model, besides the model card: a sort of changelog I want to see how & in w…

Model ReleasesDGX agent

We need a new document that AI labs should release with each new model, besides the model card: a sort of changelog I want to see how & in what way the new model changes, breaks, or improves at a rang

16 Apr 2026

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to valida…

Model ReleasesDGX agent

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to validation: finding vulnerability signal is getting cheaper; turni

21 Jul 2026

My 2hr workshop on Open vs Closed models, reward hacking, benchmaxxing & RL is out! 1. Closed vs open models 2. Throughput maxxing but accur…

Model ReleasesDGX agent

My 2hr workshop on Open vs Closed models, reward hacking, benchmaxxing & RL is out! 1. Closed vs open models 2. Throughput maxxing but accuracy minimizing 3. Benchmaxxing & cheating 4. Distillation &

Today we’re releasing Poolside Laguna S 2.1 It is a 118B-total, 8B-active open-weight model built for agentic coding and long-horizon work, …

Model ReleasesDGX agent

Today we’re releasing Poolside Laguna S 2.1 It is a 118B-total, 8B-active open-weight model built for agentic coding and long-horizon work, with context up to 1M tokens https://poolside.ai/blog/introd

3 Jun 2026

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or dist…

AgentsDGX agent

microsoft MAI tech report is a gold mine, one of the most transparent for a model at this scale. this model uses zero synthetic data or distillation from previous models. this means reasoning, agentic

← Previous
1
Next →
6,069 results
← Previous
123…102
Next →