AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,114 results
Model Releases

This is exactly why we believe in customization. Quick context: Factory's original secret scanner was deterministic, so it either flagged th…

DGX agent

This is exactly why we believe in customization. Quick context: Factory's original secret scanner was deterministic, so it either flagged things that weren't actually secrets (false positives) or miss

model-releasesfireworks-ai--x
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Claude Sonnet 5 is now available in Devin Desktop and Devin CLI. Sonnet 5 pairs frontier-level coding performance with a more affordable pri…

DGX agent

Claude Sonnet 5 has been integrated into Devin Desktop and Devin CLI, offering advanced coding capabilities at a more competitive price point than previous models. This release represents an update to

model-releasescognition-ai--x
30 Jun 2026
Model Releases

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's t…

DGX agent

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's the largest AI model theft ever attempted. The same week, the

model-releasesemad-mostaque--x
29 Jun 2026
Model Releases

Claude in Microsoft Foundry is now generally available, hosted on Azure. Azure customers get Claude Opus 4.8 and Claude Haiku 4.5, with Azur…

DGX agent

Claude models are now generally available through Microsoft Foundry on Azure infrastructure, providing Azure customers access to Claude Opus 4.8 and Claude Haiku 4.5. This integration allows enterpris

model-releasesboris-cherny--x
29 Jun 2026
Model Releases

Elon Musk turns 55 today. Here are 55 milestones. Age 54: world's first trillionaire Age 54: takes SpaceX public Age 54: SpaceX acquires xAI…

DGX agent

Elon Musk turns 55 today. Here are 55 milestones. Age 54: world's first trillionaire Age 54: takes SpaceX public Age 54: SpaceX acquires xAI Age 54: releases Grok 4 Age 53: launches Robotaxi service A

model-releaseselon-musk--x
28 Jun 2026
Model Releases

https://huggingface.co/collections/deepseek-ai/deepspec

DGX agent

https://huggingface.co/collections/deepseek-ai/deepspec Good guy DeepSeek gives us accelerated models The most interesting one here is Gemma4-12B, I presume vision included. Might be the best local mo

model-releasesclem-delangue--x
28 Jun 2026
Model Releases

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Ex…

DGX agent

Have been taking different local open-weight LLMs for a test drive in different harnesses (Qwen-Code, Codex, Claude Code). 30B Mixture-of-Expert models are kind of a nice sweet spot and can solve chal

model-releasessebastian-raschka--x
26 Jun 2026
Model Releases

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Code…

DGX agent

In a joint Fireworks and @Faros_AI evaluation of 211 real engineering tasks, Claude Code + GLM-5.2 beat both Claude Code + Opus 4.8 and Codex + GPT-5.5: - Judge score: 0.568 vs. 0.521 and 0.466 - Time

model-releasesfireworks-ai--x
25 Jun 2026
Model Releases

Big Tech's $1 trillion AI moat just got DESTROYED by a free Chinese download. Microsoft, Amazon, Google, and Meta are pouring fortunes into …

DGX agent

Big Tech's 1 trillion AI moat just got DESTROYED by a free Chinese download. Microsoft, Amazon, Google, and Meta are pouring fortunes into chips and data centers because they have been told that whoev

model-releasesgary-marcus--x
24 Jun 2026
Model Releases

btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's. now they have beat deepseek with t…

DGX agent

btw Zai IPO'ed in Jan at HK$120 a share. when I first met @louszbd nobody really knew anyone using GLM's. now they have beat deepseek with the world's undisputed top open model and in some respects (s

model-releasesswyx--x
24 Jun 2026
Industry

Building a Gitlawb Nodes adapter for Hugging Face buckets. Decentralized git storage should be able to live on the most supportive OSS infra…

DGX agent

This post discusses developing a GitLab nodes adapter to enable Hugging Face model storage on decentralized git infrastructure, leveraging open-source platforms for distributed model repository hostin

industryclem-delangue--x
24 Jun 2026
Model Releases

ParseBench is now also available on Papers with Code! Find it here: https://paperswithcode.co/benchmark/parsebench

DGX agent

ParseBench is now also available on Papers with Code! Find it here: https://paperswithcode.co/benchmark/parsebench We benchmarked Mistral OCR against other frontier and open-weight models on ParseBenc

model-releasesjerry-liu--x
24 Jun 2026
Agents

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fu…

DGX agent

How does it work? Sakana Fugu is itself an LLM, trained to call various LLMs in an agent pool, including instances of itself recursively. Fugu dynamically orchestrates the world's best models to tackl

agentsdavid-ha--x
22 Jun 2026
Model Releases

An hour in and first impression is definitely that GLM is really solid (very easy to set up on @FireworksAI_HQ, props to them for that, took…

DGX agent

A user shares positive early impressions of GLM (likely a language model), praising its solid performance and ease of setup on Fireworks AI's platform. The post highlights Fireworks AI's developer exp

model-releasesfireworks-ai--x
21 Jun 2026
Model Releases

I have some very big news... KernelBench-Hard with H100 and B200 (single gpu results) AND KernelBench-Mega tested on RTX PRO 6000, H100, B20…

DGX agent

I have some very big news... KernelBench-Hard with H100 and B200 (single gpu results) AND KernelBench-Mega tested on RTX PRO 6000, H100, B200 is finally out! Starting with Mega, each of models wrote a

model-releasesclem-delangue--x
20 Jun 2026
Model Releases

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info ne…

DGX agent

Really enjoyed reading the Microsoft MAI-Thinking-1 'Building a Hill Climbing Machine' paper. Amazing they publicly released all the info needed to train a frontier model, down to hparams. I also thou

model-releasesyann-lecun--x
10 Jun 2026
Model Releases

wooh https://x.com/shadcn/status/2064671802509410806?s=46

DGX agent

wooh https://x.com/shadcn/status/2064671802509410806?s=46 You have Claude Fable for only a few days. Here's how to make the most of it. Introducing /improve: use your most capable model to audit your

model-releasesswyx--x
10 Jun 2026
Model Releases

Introducing the Fast Gemma Challenge with Hugging Face Over the next few days, dozens of agents will collaborate to make Gemma 4 E4B even fa…

DGX agent

Google and Hugging Face are launching the Fast Gemma Challenge, where multiple agents will collaborate to optimize the performance and speed of Gemma 4 E4B model. The initiative aims to improve the ef

model-releasesclem-delangue--x
9 Jun 2026
Model Releases

Read the blog to learn more: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-live-3-5-translate/

DGX agent

This blog post from Google AI announces features and updates related to Gemini models, likely covering new capabilities for Gemini Live, version 3.5, and translation functionality. The announcement de

model-releasesgoogle-ai--x
9 Jun 2026
Model Releases

Sovereign AI for all.

DGX agent

Cohere advocates for democratizing access to sovereign AI systems, enabling organizations and nations to develop and deploy their own AI models independently rather than relying on centralized provide

model-releasescohere--x
9 Jun 2026
Model Releases

There has been a lot of hand wringing on the appropriate valuation of SpaceX. Some large institutions believe SpaceX can only be valued at h…

DGX agent

There has been a lot of hand wringing on the appropriate valuation of SpaceX. Some large institutions believe SpaceX can only be valued at half what the market seems to be willing to pay for it. Other

model-releaseselon-musk--x
9 Jun 2026
Model Releases

We've got 13 days to burn as much tokens as humanly possible on Claude Max plans Before they revert to API based billing 💀

DGX agent

We've got 13 days to burn as much tokens as humanly possible on Claude Max plans Before they revert to API based billing 💀 Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for gen

model-releasesjerry-liu--x
9 Jun 2026
Model Releases

At least until (if?) rapid improvement stops, it seems less likely someone is going to catch the Big Three AI Labs. Microsoft and Meta relea…

DGX agent

At least until (if?) rapid improvement stops, it seems less likely someone is going to catch the Big Three AI Labs. Microsoft and Meta released their models, which were fine, but not frontier. SpaceX

model-releasesethan-mollick--x
5 Jun 2026
Local Ai

https://ollama.com/library/gemma4/tags

DGX agent

Gemma4 is a language model available through Ollama's model library with multiple tagged versions for different use cases and configurations. The Ollama platform enables users to run open-source large

local-aiollama--x
5 Jun 2026
Model Releases

Nemotron 3 Ultra!

DGX agent

Nemotron 3 Ultra is an advanced language model released by NVIDIA, representing an improvement over previous versions in the Nemotron series with enhanced capabilities for various NLP tasks. The annou

model-releasesclem-delangue--x
4 Jun 2026
Model Releases

@nvidia @nebiustf More info on the new Nemotron 3 Ultra: https://x.com/NVIDIAAI/status/2062521325076299981?s=20

DGX agent

@nvidia @nebiustf More info on the new Nemotron 3 Ultra: https://x.com/NVIDIAAI/status/2062521325076299981?s=20 Today we're shipping Nemotron 3 Ultra. A 550B MoE frontier-intelligence open model built

model-releasesnous-research--x
4 Jun 2026
Model Releases

@huggingface @Gradio Registration closes tomorrow, Wednesday, June 3rd. Register here: https://huggingface.co/spaces/build-small-hackathon/r…

DGX agent

@huggingface @Gradio Registration closes tomorrow, Wednesday, June 3rd. Register here: https://huggingface.co/spaces/build-small-hackathon/registration If you're looking for a model to use to tackle t

model-releasescohere--x
2 Jun 2026
Model Releases

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use age…

DGX agent

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use agent model that beats Qwen3.5-397B, Kimi-K2.5, and Sonnet 4.6! Si

model-releasesclem-delangue--x
2 Jun 2026
Industry

Agree! was talking about this with @havoyan just a few days ago. That's also the reason why so much of the value has been accruing to the fr…

DGX agent

Agree! was talking about this with @havoyan just a few days ago. That's also the reason why so much of the value has been accruing to the frontier models in my opinion (cc @GavinSBaker) because if you

industryclem-delangue--x
29 May 2026
Model Releases

Took me a while to figure out what all the ESMFold2 rage was about. At first, the benchmarking data didn't look super remarkable to me but i…

DGX agent

Took me a while to figure out what all the ESMFold2 rage was about. At first, the benchmarking data didn't look super remarkable to me but it turns there are many impressive aspects: - Fully open sour

model-releasesyann-lecun--x
29 May 2026
Model Releases

Beyond being fast, LiteParse is designed to provide highly accurate, semantically coherent text for LLM use. We benchmarked every open-sourc…

DGX agent

Beyond being fast, LiteParse is designed to provide highly accurate, semantically coherent text for LLM use. We benchmarked every open-source, model-free PDF parser on LLM QA tasks - from PyPDF to PyM

model-releasesjerry-liu--x
28 May 2026
Model Releases

Qwen 3.7 Max is now supported in Hermes Agent

DGX agent

Nous Research has added support for Qwen 3.7 Max, a large language model, within their Hermes Agent framework. This integration enables users to leverage Qwen 3.7 Max's capabilities when building or d

model-releasesnous-research--x
26 May 2026
Industry

Product >> Brand

DGX agent

Product >> Brand 2 years ago: 'Elon destroyed the Tesla brand globally! It'll never recover!' Today: #1 selling EV in California: Tesla Model Y #1 selling EV in the U.S.: Tesla Model Y #1 selling EV i

industryelon-musk--x
25 May 2026
Model Releases

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the …

DGX agent

Shipper’s “After Automation” is the best framing I’ve read of why AI keeps generating more work, not less. The line that stuck with me: the frame is not the framer. Models climb whatever benchmark we

model-releasesitamar-friedman--x
23 May 2026
Model Releases

oops! wild update, strongly supports @emollick’s overall take:

DGX agent

oops! wild update, strongly supports @emollick’s overall take: I have to eat crow on this, in light of further information. whatever OpenAI spent on Erdos using a new model, apparently you can get GPT

model-releasesgary-marcus--x
22 May 2026
Model Releases

We are making our discount permanent! 🎉 Enjoy building with DeepSeek-V4-Pro and bring your innovative ideas to life! 🚀

DGX agent

DeepSeek has announced a permanent discount for its DeepSeek-V4-Pro model, encouraging developers to build and innovate with the platform. The announcement was made via social media and emphasizes the

model-releasesjeremy-howard--x
22 May 2026
Model Releases

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.…

DGX agent

Cursor's new Composer 2.5 takes third on the Artificial Analysis Coding Agent Index and is ~10-60x lower cost than the higher-effort Opus 4.7 and GPT-5.5 variants above it. This release puts Composer

model-releasesfireworks-ai--x
21 May 2026
Model Releases

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and …

DGX agent

brain dump of how/why we use Evals to measure agents before & after shipping to prod 1. Good Evals simulate what our real users will do and encounter. They’re not really random benchmark tasks, they r

model-releasesharrison-chase--x
20 May 2026
Model Releases

✨ Personal AI is the next computing platform. AI is shifting from something you access to something you build with, locally, at the edge, an…

DGX agent

✨ Personal AI is the next computing platform. AI is shifting from something you access to something you build with, locally, at the edge, and across systems. We’re unlocking new possibilities for deve

model-releasesclem-delangue--x
20 May 2026
Model Releases

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google…

DGX agent

Reporting from Google I/O 2026 with the four biggest themes from one of the biggest AI labs in the world. 🎤 Voice AI as an interface Google and Samsung announced new Gemini-powered glasses with Gentle

model-releasesallie-k--miller--x
20 May 2026
Model Releases

Worldlier: native support for 48 world languages and improved efficiency in non-European languages.

DGX agent

Worldlier is a Cohere language model that natively supports 48 world languages with improved efficiency, particularly for non-European languages. This represents an expansion of language coverage beyo

model-releasescohere--x
20 May 2026
Model Releases

Disappointing pricing trend with Gemini 3.5 Flash. 22.5x pricier than 2.0 Flash which came out 15 months ago (9.00 vs 0.40). Are Flash mod…

DGX agent

Disappointing pricing trend with Gemini 3.5 Flash. 22.5x pricier than 2.0 Flash which came out 15 months ago (9.00 vs 0.40). Are Flash models supposed to get this much more expensive, or is Pro just b

model-releasesjeremy-howard--x
19 May 2026
Model Releases

Gemini 3.5 Flash is now available in Windsurf!

DGX agent

Cognition AI announced the availability of Gemini 3.5 Flash within Windsurf, their AI development environment. This integration brings Google's faster Gemini 3.5 Flash model to Windsurf users, likely

model-releasescognition-ai--x
19 May 2026
Model Releases

My notes on Gemini 3.5 Flash - 3x the price of Gemini 3 Flash but Google are planning to use it for many of their own products https://simon…

DGX agent

Google's Gemini 3.5 Flash model costs approximately 3x more than Gemini 3 Flash, despite being a newer version. Google plans to integrate Gemini 3.5 Flash into many of their own products, suggesting t

model-releasessimon-willison--x
19 May 2026
Model Releases

Reference anything: Gemini Omni extends Gemini's native multimodality, allowing you to blend combinations of text, audio, image, and video i…

DGX agent

Gemini Omni is an extension of Google's Gemini model that enhances its multimodal capabilities by enabling seamless integration of text, audio, image, and video inputs and outputs. This advancement al

model-releasesgoogle-ai--x
19 May 2026
Model Releases

The table in HTML format for easier (and non-truncated) viewing: https://sebastianraschka.com/llm-architecture-gallery/active-parameter-rati…

DGX agent

Sebastian Raschka shared an HTML-formatted table comparing active parameter ratios across different large language model architectures for improved readability and to avoid text truncation. The resour

model-releasessebastian-raschka--x
14 May 2026
Model Releases

@BereznevKi20669 @ggerganov Yes I believe the real llama.cpp revolution is yet to happen at its full scale. As computers will have more RAM …

DGX agent

@BereznevKi20669 @ggerganov Yes I believe the real llama.cpp revolution is yet to happen at its full scale. As computers will have more RAM and models will improve, and *if* China will continue shippi

model-releasesgeorgi-gerganov--x
12 May 2026
Model Releases

When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last com…

DGX agent

When is the last time a general purpose LLM (putting aside hybrid systems like Claude Code with special purpose symbolic harnesses) last completely blew away all competing prior models? GPT-4 relative

model-releasesgary-marcus--x
12 May 2026
← Previous
1…4041424344…128
Next →