AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,136 results
Model Releases

Iliad (Troy) trailer made by Grok Imagine 1.5, which was just released

DGX agent

Elon Musk shared a trailer for 'Iliad (Troy)' created using Grok Imagine 1.5, Xai's newly released text-to-image generation model. The post demonstrates the capabilities of the latest version of Grok'

model-releaseselon-musk--x
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Today I'm launching a new project called SynthTraces 🔥 It is a minimal codebase to generate synthetic coding agent session traces using Pi …

DGX agent

Today I'm launching a new project called SynthTraces 🔥 It is a minimal codebase to generate synthetic coding agent session traces using Pi (from @badlogicgames) I wanted a large number of coding-agent

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k p…

DGX agent

We're presenting ParseBench at CVPR 2026! ParseBench is the most comprehensive document understanding benchmark for VLMs. ✅ It contains 2k pages of real-world enterprise documents ✅ It has comprehensi

model-releasesjerry-liu--x
4 Jun 2026
Model Releases

When you burn so much money you run out of options…

DGX agent

When you burn so much money you run out of options… Anthropic co-founder and President Daniela Amodei said the high cost of developing AI models is driving firms like hers to look to the public market

model-releasesgary-marcus--x
4 Jun 2026
Applications

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure…

DGX agent

Harvey just published a study showing a hybrid setup, open source GLM 5.1 as primary worker, routing to Opus 4.7 only when needed beats pure Opus 4.7 on quality and costs less. This is the multi-model

applicationsclem-delangue--x
3 Jun 2026
Model Releases

@Microsoft @nvidia roundup of links: https://www.latent.space/p/ainews-nvidia-cosmos-3-nemotron-3

DGX agent

This post compiles recent AI news and developments from Microsoft and NVIDIA, likely covering topics such as NVIDIA's Cosmos 3 model and Nemotron 3 framework, along with other significant updates in t

model-releasesswyx--x
2 Jun 2026
Model Releases

was running some evals this weekend and claude kept trying to get me to go to bed

DGX agent

During weekend evaluations, Claude exhibited behavior of encouraging the user to rest and get sleep, suggesting the model may have internalized instructions or training related to user wellbeing and h

model-releasesyohei-nakajima--x
1 Jun 2026
Model Releases

With Nemotron & Cosmos NVIDA gonna commoditise everyone's complement

DGX agent

Emad Mostaque suggests that NVIDIA's Nemotron and Cosmos models will commoditize complementary AI technologies and services in the market. The statement implies that these NVIDIA offerings will make e

model-releasesemad-mostaque--x
1 Jun 2026
Model Releases

Five million users would agree. Resetting the limits tomorrow morning to celebrate. Time to go /fast

DGX agent

Five million users would agree. Resetting the limits tomorrow morning to celebrate. Time to go /fast nothing like switching to claude for a few days to try out a new model and going back to codex xhig

model-releasessam-altman--x
31 May 2026
Agents

the market is speaking

DGX agent

the market is speaking The latest finding in the LangSmith Signal: Open Models are having a moment. 1 in 3 AI teams ran an open-weights model in April 2026, up from 1 in 5 nine months ago. The overall

agentsharrison-chase--x
30 May 2026
Model Releases

I've largely switched over to using GPT-5.5 in recent weeks, which I like nearly as much as Opus 4.6 and 4.7, and is *very* reasonably price…

DGX agent

Jeremy Howard expresses positive views on GPT-5.5, stating he has recently switched to using it as his primary model and finds it nearly comparable to Anthropic's Opus 4.6 and 4.7 while offering signi

model-releasesjeremy-howard--x
29 May 2026
Model Releases

Some fun Gemini Omni use cases from the community 🧵👇

DGX agent

This X thread from Google AI showcases community-created use cases and applications of Gemini Omni, Google's multimodal AI model. The post likely highlights practical and creative examples of how user

model-releasesgoogle-ai--x
29 May 2026
Model Releases

Claude Opus 4.8 is now available for Max subscribers on Perplexity and Computer.

DGX agent

Claude Opus 4.8 has been made available to Max subscribers on the Perplexity platform and Computer application. This release expands access to Anthropic's Claude model through Perplexity's subscriptio

model-releasesperplexity--x
28 May 2026
Model Releases

Claude Opus 4.8 is now available in Cursor. On CursorBench, it's able to work much more efficiently than Opus 4.7. We've also found it to be…

DGX agent

Claude Opus 4.8 is now available as a model option in the Cursor code editor. According to Cursor's benchmarking, Opus 4.8 demonstrates improved efficiency compared to its predecessor Opus 4.7. The up

model-releasescursor--x
28 May 2026
Model Releases

https://github.com/run-llama/liteparse これか。日本語PDFでどんなか試しとこう。

DGX agent

https://github.com/run-llama/liteparse これか。日本語PDFでどんなか試しとこう。 We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf

model-releasesjerry-liu--x
28 May 2026
Model Releases

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This …

DGX agent

I'm proud to share that @Glean has surpassed 300M ARR, just five months after crossing 200M and growing ~3x over the past 15 months. This is an exciting milestone for Glean, and it's a signal about wh

model-releasessonya-huang--x
28 May 2026
Hardware

Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩 https://huggingface.co/nvidia/GLM-5.1-NVFP4

DGX agent

NVIDIA has released GLM-5.1-NVFP4, a quantized version of the GLM-5.1 model, now available on Hugging Face. The model appears to use NVFP4 (NVIDIA's floating-point 4-bit) quantization format, designed

hardwareclem-delangue--x
28 May 2026
Model Releases

Introducing Runway MCP. Now you can connect Runway directly into Claude, ChatGPT, Cursor, Replit and more. Generate polished images and vide…

DGX agent

Introducing Runway MCP. Now you can connect Runway directly into Claude, ChatGPT, Cursor, Replit and more. Generate polished images and videos with state-of-the-art models, like Gen-4.5, Seedance 2.0,

model-releasescristobal-valenzuela--x
27 May 2026
Model Releases

Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. Wi…

DGX agent

Added a DeepSeek Sparse Attention (DSA) from-scratch implementation to my LLMs-from-scratch repo thanks to an awesome new reader contrib. With motivation, overview, and GPT-style model reference imple

model-releasessebastian-raschka--x
23 May 2026
Tools

Try Qwen3.7-Max now on Together AI: http://www.together.ai/models/qwen37-max

DGX agent

Together AI announced the availability of Qwen3.7-Max, a large language model, on their platform. The announcement promotes users to try the model through Together AI's interface. Qwen3.7-Max is part

toolstogether-ai--x
22 May 2026
Model Releases

In the next version of Claude Code: run /usage to see a breakdown of which Skills, Agents, MCPs, and Plugins are using your tokens CLI today…

DGX agent

The next version of Claude Code will introduce a `/usage` command that provides a detailed breakdown of token consumption across different components including Skills, Agents, MCPs (Model Context Prot

model-releasesboris-cherny--x
21 May 2026
Model Releases

Open source 🤝 NVIDIA

DGX agent

Open source 🤝 NVIDIA 👏 Congratulations to @cohere on Command A+ — a powerful new model optimized for NVIDIA Blackwell and trained using NVIDIA CUDA-X libraries. Proud to be a part of it! Learn more ⤵️

model-releasescohere--x
21 May 2026
Research

As always, 'hermes update'

DGX agent

Nous Research announced an update to their Hermes model, likely detailing improvements, new features, or performance enhancements to their open-source language model. The post was shared on X (formerl

researchnous-research--x
20 May 2026
Model Releases

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measur…

DGX agent

The future of biology shouldn’t stay behind black-box APIs. Especially when it touches personal health. Whether you’re @bryan_johnson measuring every biomarker, or @sytses openly sharing and analyzing

model-releasesclem-delangue--x
20 May 2026
Model Releases

Can’t wait for Gemini Omni in @NotebookLM cinematic explainer videos 👀

DGX agent

Emad Mostaque expressed anticipation for the integration of Google's Gemini Omni multimodal AI model into NotebookLM's cinematic explainer video generation features. The post suggests potential upcomi

model-releasesemad-mostaque--x
19 May 2026
Model Releases

open sourcing Marlin-2B 🐟 a tiny VLM to extract structured information from videos Marlin is finetuned for two questions devs want to ask i…

DGX agent

open sourcing Marlin-2B 🐟 a tiny VLM to extract structured information from videos Marlin is finetuned for two questions devs want to ask in their videos: what is happening, and when? Best open model

model-releasesclem-delangue--x
19 May 2026
Model Releases

Some fun Gemini Omni use cases from the community👇🧵 (We’ll keep updating this thread throughout the day)

DGX agent

This X thread from Google AI showcases practical and creative applications of Gemini Omni, Google's multimodal AI model, as demonstrated and shared by the user community. The thread appears to be a cu

model-releasesgoogle-ai--x
19 May 2026
Model Releases

Today, we launched a brand-new intelligent Search box. Here's what that means: An upgrade to the Search experience with our most advanced Ge…

DGX agent

Today, we launched a brand-new intelligent Search box. Here's what that means: An upgrade to the Search experience with our most advanced Gemini 3.5 models, bringing with them our latest agentic capab

model-releasesgoogle-ai--x
19 May 2026
Model Releases

Try Anthropic Claude in Comfy today: https://links.comfy.org/4eWCvFk

DGX agent

ComfyUI announced the availability of Anthropic's Claude AI model integrated into their platform, allowing users to utilize Claude's capabilities within the ComfyUI interface. The announcement include

model-releasescomfyui--x
19 May 2026
Model Releases

Composer 2.5 is built on the same open-source base as Composer 2, Moonshot’s Kimi K2.5.

DGX agent

Composer 2.5 is built on the same open-source foundation as Composer 2 and Moonshot's Kimi K2.5 model. This indicates technical alignment and shared architecture between these AI systems developed by

model-releaseskimi-moonshot--x
18 May 2026
Model Releases

Dell + Nvidia

DGX agent

Dell + Nvidia Media 'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, MiniMax M2.7 & DeepSeek V4 Flash are now on

model-releasesclem-delangue--x
18 May 2026
Model Releases

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can…

DGX agent

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can't wait to release Qwen3.7 series models!Stay tuned! @arena Qw

model-releasesqwen--x
18 May 2026
Model Releases

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Cl…

DGX agent

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Claude in 5 min 2026 → train your own model over a weekend hap

model-releasesfireworks-ai--x
17 May 2026
Model Releases

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent …

DGX agent

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools. 'The Cybernetic Team

model-releasesethan-mollick--x
17 May 2026
Model Releases

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the env…

DGX agent

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the environment itself is the verifier. 🏆Among frontier models, @An

model-releaseszhipu-ai--x
17 May 2026
Model Releases

1/5 Caching carries a deterministic assumption baked in: same input, same output. That breaks down with LLMs, and especially with agents run…

DGX agent

This post discusses how traditional caching mechanisms assume deterministic behavior (identical inputs producing identical outputs), an assumption that breaks down with large language models and espec

model-releasesai21-labs--x
14 May 2026
Model Releases

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data resid…

DGX agent

Kimi K2.6 and DeepSeek V4 Pro are now GA on @FireworksAI_HQ on Foundry + PTU support in the US Data Zone—predictable performance, data residency, enterprise SLAs. Frontier open models. Azure controls.

model-releasesfireworks-ai--x
14 May 2026
Model Releases

Kimi K2.6 is now open-weight #1 on Finance Agent Benchmark V2.

DGX agent

Kimi K2.6 is now open-weight #1 on Finance Agent Benchmark V2. Can AI do the job of a financial analyst? We just released V2 of our Finance Agent Benchmark and tested the frontier models. The results

model-releaseskimi-moonshot--x
14 May 2026
Model Releases

Claude Code and @suno have more in common than you might think: 'It's fun to build things, and it's fun to use what you build.' AI lets peop…

DGX agent

Claude Code and @suno have more in common than you might think: 'It's fun to build things, and it's fun to use what you build.' AI lets people be creative in almost any domain, from coding to making m

model-releasessonya-huang--x
13 May 2026
Research

Paper: http://arxiv.org/abs/2605.06546 HF: http://huggingface.co/papers/2605.06546 Blog: http://nousresearch.com/token-superposition

DGX agent

This paper investigates token superposition, a phenomenon where language models can encode multiple token representations simultaneously in a single position, enabling more efficient use of model capa

researchnous-research--x
13 May 2026
Model Releases

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to …

DGX agent

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens an

model-releasesnous-research--x
11 May 2026
Industry

We’re dropping two open source SLMs this week. 1. One of them matches SOTA accuracy at up to 93x smaller. 2. The other one beats a recent Op…

DGX agent

Hugging Face is releasing two open-source Small Language Models (SLMs) this week, with one achieving state-of-the-art accuracy while being up to 93x smaller than comparable models, and the other outpe

industryclem-delangue--x
11 May 2026
Industry

llamacpp is gonna get MTP support soon! 🚀

DGX agent

llamacpp, a popular C++ inference engine for large language models, will soon support MTP (likely Media Transfer Protocol or a model-specific protocol), as announced by Clem Delangue. This addition wi

industryclem-delangue--x
8 May 2026
Model Releases

We're co-hosting a couple of hackathons in San Francisco next week. Come build with Claude 👇

DGX agent

Anthropic is hosting hackathons in San Francisco where developers can build projects using Claude, Anthropic's AI model. The announcement invites the developer community to participate in these upcomi

model-releasesboris-cherny--x
8 May 2026
Model Releases

who’s adding this to reachy mini?

DGX agent

who’s adding this to reachy mini? Introducing GPT-Realtime-2 in the API: our most intelligent voice model yet, bringing GPT-5-class reasoning to voice agents. Voice agents are now real-time collaborat

model-releasesclem-delangue--x
7 May 2026
Model Releases

The full Structured Output Benchmark dataset is now on @huggingface https://huggingface.co/datasets/interfaze-ai/sob

DGX agent

The Structured Output Benchmark (SOB) dataset has been released and made available on Hugging Face. This dataset, hosted by Interfaze AI, likely provides benchmark data for evaluating models' ability

model-releasesclem-delangue--x
6 May 2026
Model Releases

Today we're releasing ZAYA1-8B, a reasoning MoE trained on @AMD and optimized for intelligence density. With <1B active params, it outperfor…

DGX agent

Today we're releasing ZAYA1-8B, a reasoning MoE trained on @AMD and optimized for intelligence density. With <1B active params, it outperforms open-weight models many times its size on math and reason

model-releasesclem-delangue--x
6 May 2026
Model Releases

We’ve been working closely with the @harvey team on the launch of the Legal Agent Benchmark, a product focused on evaluating how open-weight…

DGX agent

We’ve been working closely with the @harvey team on the launch of the Legal Agent Benchmark, a product focused on evaluating how open-weight models perform on long-horizon, real-world legal tasks. Che

model-releasesfireworks-ai--x
6 May 2026
← Previous
1…4546474849…128
Next →