AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,102 results
Model Releases

That's a badass title, and it's true! Every day, it gets harder and harder to create tests that AI models can't beat. Reality is humanity's …

DGX agent

That's a badass title, and it's true! Every day, it gets harder and harder to create tests that AI models can't beat. Reality is humanity's real last exam. Andon Labs' Real-World AI Evals: Claude call

model-releasesswyx--x
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reache…

DGX agent

Frontier models are powerful advisors. On @harvey's Legal Agent Benchmark, a GLM 5.1 worker using Claude Opus 4.7 as a sparse advisor reached 18/100 all-pass versus 14/100 for Opus alone, at 39% of th

model-releasesfireworks-ai--x
3 Jun 2026
Model Releases

Gemma 4 model load issues fixed in engine version 2.20.1. lms runtime update --all

DGX agent

Gemma 4 model load issues fixed in engine version 2.20.1. lms runtime update --all Gemma 4 12B is here! Dense, mid-sized Gemma that fits right on your laptop - released by @google under Apache 2.0 Ava

model-releaseslm-studio--x
3 Jun 2026
Local Ai

Model page: https://ollama.com/library/gemma4

DGX agent

Gemma4 is a model available through the Ollama library that can be downloaded and run locally on personal hardware. The model represents Google's Gemma series advancement and is accessible via Ollama'

local-aiollama--x
3 Jun 2026
Model Releases

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + …

DGX agent

Startup discovery platform @harmonic_ai rebuilt Scout, their AI platform using Deep Agents and LangSmith. Deep Agents: One frontier model + two tool sets (global company data and firm-specific context

model-releasesharrison-chase--x
3 Jun 2026
Model Releases

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-…

DGX agent

We’re bringing new capabilities to GPT-Rosalind, a model series purpose-built for life sciences research at enterprise scale. It brings GPT-5.5’s agentic coding and tool use together with stronger int

model-releasesopenai--x
3 Jun 2026
Industry

Step-3.7-Flash from @StepFun_ai is a silent winner. Super impressive results, the best model under 500B params on HF leaderboards.. All whil…

DGX agent

Step-3.7-Flash is a compact language model from StepFun AI that reportedly achieves top-tier performance among models under 500 billion parameters on Hugging Face leaderboards, despite receiving limit

industryclem-delangue--x
2 Jun 2026
Model Releases

We're sponsoring a hackathon to scale down. Hosted by our friends @huggingface and @Gradio, we want working with models to feel like yours a…

DGX agent

We're sponsoring a hackathon to scale down. Hosted by our friends @huggingface and @Gradio, we want working with models to feel like yours again. Small enough that it's inexpensive to run, big enough

model-releasescohere--x
2 Jun 2026
Model Releases

In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generatio…

DGX agent

In May, we integrated 11 new models spanning image, 3D, audio, video, and multimodal. The highlights: → Krea 2 — style-first image generation, live as a Partner Node on day one. Competes on how the fr

model-releasescomfyui--x
1 Jun 2026
Model Releases

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

model-releasesfireworks-ai--x
1 Jun 2026
Model Releases

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI thr…

DGX agent

OpenAI frontier models and Codex are now generally available on AWS, giving enterprises a new way to build on Amazon Bedrock with OpenAI through the security, compliance, and governance workflows they

model-releasesopenai--x
1 Jun 2026
Model Releases

We need more coding and agent traces public sharing to build datasets and better open source models! Lots of people contributing already, yo…

DGX agent

We need more coding and agent traces public sharing to build datasets and better open source models! Lots of people contributing already, you should share yours too! https://huggingface.co/datasets?se

model-releasesclem-delangue--x
31 May 2026
Safety

is having a four month lead a sustainable multitrillion dollar business model?

DGX agent

is having a four month lead a sustainable multitrillion dollar business model? We took another look at the capability gap between open-weight and proprietary models. Since the start of the year, open-

safetygary-marcus--x
30 May 2026
Model Releases

models underestimate how much work it takes (token usage) to accomplish a task, just like us

DGX agent

models underestimate how much work it takes (token usage) to accomplish a task, just like us 🧵 Claude-Opus-4.8 takes you too much tokens - but is this issue general across agents? Do agents know how m

model-releasesyohei-nakajima--x
29 May 2026
Model Releases

Claude Opus 4.8 is out today. It's our strongest coding model yet: up on SWE-bench Pro (from 64.3 to 69.2) and noticeably more honest about …

DGX agent

Claude Opus 4.8 is out today. It's our strongest coding model yet: up on SWE-bench Pro (from 64.3 to 69.2) and noticeably more honest about its own work. It tells you when it's unsure and catches its

model-releasesboris-cherny--x
28 May 2026
Tools

This tracks. 30 trillion tokens a day on our end, and open model share keeps climbing. Our partners @FactoryAI are seeing what we're seeing.

DGX agent

This tracks. 30 trillion tokens a day on our end, and open model share keeps climbing. Our partners @FactoryAI are seeing what we're seeing. Narrative violation: Open model use in Factory has more tha

toolsfireworks-ai--x
28 May 2026
Model Releases

Krea is now built in to Hermes Agent as an image generation API provider, allowing your agent to use Krea 2: a new foundation model trained …

DGX agent

Krea is now built in to Hermes Agent as an image generation API provider, allowing your agent to use Krea 2: a new foundation model trained from scratch to balance aesthetic quality and fine control,

model-releasesnous-research--x
27 May 2026
Model Releases

We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf,…

DGX agent

We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf, pypdf, markitdown, pdftotext, opendataloader, pymupdf4llm)

model-releasesjerry-liu--x
27 May 2026
Industry

🙏 Thank you all for the incredible love and support! Our latest Tencent Hunyuan translation models are on fire on Hugging Face: 🥰Hy-MT2-1.…

DGX agent

🙏 Thank you all for the incredible love and support! Our latest Tencent Hunyuan translation models are on fire on Hugging Face: 🥰Hy-MT2-1.8B ranks #1 🥰Hy-MT2-30B-A3B ranks #4 on the open-source model

industryclem-delangue--x
26 May 2026
Industry

Today we’re releasing 1-bit and Ternary Bonsai Image 4B. A new family of image-generation models designed to run high-quality diffusion infe…

DGX agent

Hugging Face has released 1-bit and Ternary Bonsai Image 4B, a new family of lightweight image-generation models optimized for efficient diffusion inference. These models are designed to deliver high-

industryclem-delangue--x
26 May 2026
Local Ai

Try out the models Below 👇 Nanobana Pro: https://links.comfy.org/4f5k5lN GPT Image 2: https://links.comfy.org/4dAi4Nq

DGX agent

ComfyUI promoted two AI models available for testing: Nanobana Pro and GPT Image 2, providing direct links for users to access and try out these models through their platform. This appears to be a soc

local-aicomfyui--x
26 May 2026
Model Releases

I’ve just released MiMo V2.5-Coder. If you have 128 GB of RAM, this is one of the best models you can run locally. It’s fast, and in all my …

DGX agent

I’ve just released MiMo V2.5-Coder. If you have 128 GB of RAM, this is one of the best models you can run locally. It’s fast, and in all my experiments it outperformed Qwen 3.6 and DeepSeek 4-Flash. h

model-releasesclem-delangue--x
25 May 2026
Model Releases

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up o…

DGX agent

Just spent a full coding session with Grok Build and honestly? It's right there with Claude. The model is sharp, the agentic flow holds up on complex tasks, and it has actual personality. Few rough ed

model-releaseselon-musk--x
25 May 2026
Model Releases

Gemini Flash 3.5 is such a disappointing model. It's intelligence and speed is awesome. Absolutely amazing. But it's been trained to max eva…

DGX agent

Gemini Flash 3.5 is such a disappointing model. It's intelligence and speed is awesome. Absolutely amazing. But it's been trained to max evals, not to be helpful to humans. It goes off and does random

model-releasesjeremy-howard--x
22 May 2026
Model Releases

Introducing Qwen3.7-Max from @Alibaba_Qwen, Qwen’s flagship model for the agent era with 1M context and leading performance across agentic c…

DGX agent

Introducing Qwen3.7-Max from @Alibaba_Qwen, Qwen’s flagship model for the agent era with 1M context and leading performance across agentic coding, reasoning, and long-horizon autonomy. AI natives can

model-releasestogether-ai--x
22 May 2026
Safety

update @thestalwart found more recent models less vulnerable. would be good to do a broad study of this.

DGX agent

Gary Marcus notes that researcher @thestalwart has found more recent AI models to be less vulnerable to certain attacks or exploits, and suggests that a comprehensive study across multiple models woul

safetygary-marcus--x
22 May 2026
Model Releases

Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else t…

DGX agent

Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else the same. Available inside our new Edit Studio, you can work

model-releasescristobal-valenzuela--x
22 May 2026
Model Releases

New inflection point in the accelerating growth of open-source models usage is coming

DGX agent

New inflection point in the accelerating growth of open-source models usage is coming 🦔Microsoft canceled its internal Claude Code licenses this week after token-based billing made the cost untenable,

model-releasesclem-delangue--x
21 May 2026
Model Releases

Self-Evolving in the Wild:Over the course of ~35 hours of continuous autonomous execution, the model performed 432 kernel evaluations across…

DGX agent

Self-Evolving in the Wild:Over the course of ~35 hours of continuous autonomous execution, the model performed 432 kernel evaluations across 1,158 tool calls. It wrote, compiled, profiled, and iterati

model-releasesqwen--x
21 May 2026
Model Releases

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefi…

DGX agent

Today we release a study on decoupling the benefits of subword tokenization for language model training, by simulating each suspected benefit one at a time inside a 1.7B byte-level pretraining pipelin

model-releasesnous-research--x
21 May 2026
Model Releases

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular,…

DGX agent

The proof came from a general-purpose reasoning model, not a system built specifically to solve math problems or this problem in particular, and represents an important milestone for the math and AI c

model-releasesopenai--x
20 May 2026
Industry

Today, we’re sharing that a general-purpose internal @openai model achieved a breakthrough on one of the best-known combinatorial geometry p…

DGX agent

Today, we’re sharing that a general-purpose internal @openai model achieved a breakthrough on one of the best-known combinatorial geometry problems. Less than 1 year ago frontier AI models were at IMO

industrysam-altman--x
20 May 2026
Model Releases

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗

DGX agent

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗 Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it t

model-releasesclem-delangue--x
20 May 2026
Tutorials

deepagents v0.6 is about performance the first level at which we can control that is the model layer: how can you squeeze perf out of a mode…

DGX agent

deepagents v0.6 is about performance the first level at which we can control that is the model layer: how can you squeeze perf out of a model? tweaking prompts, tool names, and tool descriptions in ac

tutorialsharrison-chase--x
19 May 2026
Local Ai

Use your LM Studio models to code locally in @zeddotdev 🚀

DGX agent

Use your LM Studio models to code locally in @zeddotdev 🚀 Local model usage grew 3x in Zed's agent in the last 10 weeks. Cameron Mcloughlin on why he prefers local: 'I worry about over-reliance on pro

local-ailm-studio--x
19 May 2026
Model Releases

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories…

DGX agent

We were able to sit down with the @GoogleDeepmind team behind the new Gemini Omni Flash model to hear all of their behind-the-scenes stories, memorable moments, and many, many (occasionally embarrassi

model-releasesgoogle-ai--x
19 May 2026
Model Releases

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s →…

DGX agent

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s → 45 tok/s (+78%). Two flags on llama-server: --spec-type draf

model-releasesclem-delangue--x
18 May 2026
Model Releases

One thing to watch for with Claude & GPT is that the models expose too much irrelevant history in their outputs. Slides are given footers sa…

DGX agent

One thing to watch for with Claude & GPT is that the models expose too much irrelevant history in their outputs. Slides are given footers saying things like 'Better, more targeted version' if you aske

model-releasesethan-mollick--x
18 May 2026
Model Releases

We causally trained a lot of SOTA search models internally, shall we make some small release from time to time 🤣🤣

DGX agent

We causally trained a lot of SOTA search models internally, shall we make some small release from time to time 🤣🤣 @bo_wangbo stealth releasing probably the strongest open multilingual ColBERT (and it'

model-releasesclem-delangue--x
18 May 2026
Model Releases

'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, Mi…

DGX agent

'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, MiniMax M2.7 & DeepSeek V4 Flash are now one click away on Dell

model-releasesclem-delangue--x
18 May 2026
Industry

yeah that's pretty good xAI might be able to cook with Cursor data + 10T model

DGX agent

yeah that's pretty good xAI might be able to cook with Cursor data + 10T model Introducing Composer 2.5, our most powerful model yet. It's more intelligent, better at sustained work on long-running ta

industryelon-musk--x
18 May 2026
Model Releases

Weekends are for vibe coding. But are your vibes continuously improving? Fine-tune your own model → stop waiting on someone else's release c…

DGX agent

Weekends are for vibe coding. But are your vibes continuously improving? Fine-tune your own model → stop waiting on someone else's release cycle. Today's training update: Gemma 4 Dense is now availabl

model-releasesfireworks-ai--x
15 May 2026
Model Releases

any time a model router company drops data, its worth browsing. here we learn that gemini leads in education and personal assistants (?!), a…

DGX agent

any time a model router company drops data, its worth browsing. here we learn that gemini leads in education and personal assistants (?!), ant leads in vibecoding and koding and back office (?!), and

model-releasesswyx--x
14 May 2026
Model Releases

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in …

DGX agent

Like any AI dev not employed by a closed lab, we share the ambition that at least 10x more should be capable of training frontier models in 2026. And like we all know, some 10x leaps can't survive a c

model-releasesfireworks-ai--x
14 May 2026
Model Releases

Built a local coding harness powered by Gemma 4. It runs locally, connects to my model backend, starts coding sessions, streams responses, a…

DGX agent

Built a local coding harness powered by Gemma 4. It runs locally, connects to my model backend, starts coding sessions, streams responses, and uses tools through a CLI-style workflow. Still early, but

model-releasesollama--x
13 May 2026
Research

@SakanaAILabs @NVIDIAAI Sparser, Faster, Lighter Transformer Language Models https://arxiv.org/abs/2603.23198

DGX agent

This research paper from Sakana AI and NVIDIA explores techniques for creating more efficient transformer language models by reducing sparsity, computational requirements, and model size while maintai

researchdavid-ha--x
13 May 2026
Model Releases

LiteParse is the best open-source, model-free document parser for AI agents. Run it over over 50+ document types, and it will parse dense pa…

DGX agent

LiteParse is the best open-source, model-free document parser for AI agents. Run it over over 50+ document types, and it will parse dense pages with complex text layouts and tables, and it will extrac

model-releasesjerry-liu--x
12 May 2026
Model Releases

This NVIDIA remains the strongest platform for large-model inference at scale. Prefill/decode disaggregation, Blackwell-native quantization,…

DGX agent

This NVIDIA remains the strongest platform for large-model inference at scale. Prefill/decode disaggregation, Blackwell-native quantization, custom kernels, and rack-scale NVLink turn GB200 into faste

model-releasesperplexity--x
12 May 2026
← Previous
1…1718192021…128
Next →