AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,106 results
Tools

Open model usage has gone from 10% of AI tokens to 30% in a year. The shift to open, modular AI is here to stay. Our founders on what's driv…

DGX agent

Open source AI model usage has tripled from 10% to 30% of total AI tokens consumed within a year, reflecting a significant market shift toward open and modular AI architectures. This trend indicates g

toolstogether-ai--x
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tools

The most interesting Fable tip I've heard so far is to let the model use its own judgement as much as possible I told it 'For all coding tas…

DGX agent

The most interesting Fable tip I've heard so far is to let the model use its own judgement as much as possible I told it 'For all coding tasks use your judgement to decide an appropriate lower power m

toolssimon-willison--x
3 Jul 2026
Tools

@DecagonAI @AshwinSreenivas Under the hood: 6x cost reduction per turn, p95 latency under 400ms, and models shipping weekly. https://www.tog…

DGX agent

Together AI announced significant performance improvements in their AI infrastructure, achieving a 6x cost reduction per turn while maintaining p95 latency under 400ms, with new models being released

toolstogether-ai--x
2 Jul 2026
Tools

Our CEO @vipulved on @CNBC with @dee_bosa: your data is your recipe. As models get smarter, sending proprietary workflows, customer context,…

DGX agent

Our CEO @vipulved on @CNBC with @dee_bosa: your data is your recipe. As models get smarter, sending proprietary workflows, customer context, and business logic into closed systems becomes a strategic

toolstogether-ai--x
2 Jul 2026
Industry

Here's Pivot 0.5, our first trained model! ✨ Try our research preview for free already. Open weights coming to @huggingface soon! 🤗

DGX agent

Pivot 0.5 is the first trained model released by the organization, available as a free research preview with open weights coming soon to Hugging Face. The announcement was made by Clem Delangue on X (

industryclem-delangue--x
1 Jul 2026
Research

Can regularization based JEPA (e.g. SIGReg) scale and compete with SOTA foundation models (DINO)? Here is the answer: yes and with 10x less …

DGX agent

Can regularization based JEPA (e.g. SIGReg) scale and compete with SOTA foundation models (DINO)? Here is the answer: yes and with 10x less data. VISReg (slight variation of SIGReg) competes with DINO

researchyann-lecun--x
30 Jun 2026
Applications

Recommended reading if you are looking to scale with open models in production.

DGX agent

This post likely recommends resources or reading materials for practitioners looking to deploy and scale open-source language models in production environments. DAIR.AI, a research organization focuse

applicationsdair-ai--x
30 Jun 2026
Tools

See our full model rankings: http://cursor.com/evals

DGX agent

Cursor published a comprehensive ranking of AI models, likely evaluating them across performance metrics, capabilities, and use cases relevant to their code editor platform. The rankings help users un

toolscursor--x
30 Jun 2026
Tools

We heard your feedback. You want to go faster. Introducing GLM 5.2 Fast The same model and quality as GLM 5.2 standard, now at 140 tok/s Fli…

DGX agent

Fireworks AI announced GLM 5.2 Fast, an optimized version of their GLM 5.2 model that maintains the same quality and capabilities as the standard version while delivering significantly faster inferenc

toolsfireworks-ai--x
30 Jun 2026
Agents

This is a neat idea around doing model routing and sub-agent delegation while ensuring cache hits on accumulated context for all agents. It …

DGX agent

This is a neat idea around doing model routing and sub-agent delegation while ensuring cache hits on accumulated context for all agents. It makes a lot of sense: you want to ensure that all subagents

agentsjerry-liu--x
29 Jun 2026
Applications

Elon Musk sitting by the Model 3 production line in 2017 on his birthday Never give up 🙌

DGX agent

In 2017, Elon Musk was photographed at the Tesla Model 3 production line on his birthday, with the accompanying message 'Never give up' emphasizing persistence and determination. The image captured a

applicationselon-musk--x
28 Jun 2026
Applications

Getting regulated by a government because your model is 'too dangerous' is the best marketing (especially for enterprise sales) so everyone …

DGX agent

Clem Delangue suggests that government regulation of AI models perceived as 'dangerous' can paradoxically serve as effective marketing, particularly for enterprise sales. The implication is that regul

applicationsclem-delangue--x
28 Jun 2026
Applications

In my experience, all model routers underestimate the difficulty of non-math/coding tasks and assign them too little intelligence. This is w…

DGX agent

In my experience, all model routers underestimate the difficulty of non-math/coding tasks and assign them too little intelligence. This is worth addressing, as non-verifiable tasks (innovation, market

applicationsethan-mollick--x
28 Jun 2026
Applications

So what model is OpenAI saving the GPT-6 label for?

DGX agent

Ethan Mollick discusses OpenAI's naming convention and model release strategy, speculating on what capabilities or timeline the company might be reserving the GPT-6 designation for rather than applyin

applicationsethan-mollick--x
28 Jun 2026
Tutorials

the future of AI is multi-model (including a majority of open-source ones provided by @huggingface of course!)!

DGX agent

the future of AI is multi-model (including a majority of open-source ones provided by @huggingface of course!)! How to keep AI spend flat while token usage grows exponentially: Not with friction and s

tutorialsclem-delangue--x
28 Jun 2026
Applications

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A lea…

DGX agent

GLM 5.2 is the open coding model everyone's been talking about. Now you can fine-tune it on Fireworks. SFT, DPO, and RL all supported. A leaderboard winner can still lose on your codebase. Training cl

applicationsfireworks-ai--x
24 Jun 2026
Model Releases

Introducing Claude for Music. You can now create songs from Claude Code, Hermes, Codex, or any agent you’re using. SOTA music model @MiniMax…

DGX agent

I cannot provide an accurate summary for this entry. The URL and source attribution appear inconsistent (title credits Yohei Nakajima but URL references a different user), and the post references prod

model-releasesyohei-nakajima--x
24 Jun 2026
Agents

Unlimited OCR is a great model on table parsing and understanding proper reading order. However it does struggle a little on semantic format…

DGX agent

Unlimited OCR is a great model on table parsing and understanding proper reading order. However it does struggle a little on semantic formatting, charts (it does decent at bounding boxes). Attaching t

agentsjerry-liu--x
24 Jun 2026
Local Ai

Read the technical paper on Krea 2 https://www.krea.ai/blog/krea-2-technical-report Download the model weights https://github.com/krea-ai/kr…

DGX agent

Krea 2 is a technical advancement in AI image generation with newly released model weights available for download on GitHub. The technical report details the improvements and capabilities of this vers

local-aicomfyui--x
23 Jun 2026
Industry

Going to cross 3M public models & 1M public datasets on @huggingface in a few days. Open-source AI is on fire!

DGX agent

Hugging Face was approaching milestones of 3 million public models and 1 million public datasets on its platform, reflecting rapid growth in open-source AI resources. The announcement highlights the e

industryclem-delangue--x
22 Jun 2026
Industry

Is there a market where folk are predicting when Reflection AI will drop their first model This is probably as much compute as currently use…

DGX agent

Is there a market where folk are predicting when Reflection AI will drop their first model This is probably as much compute as currently used by all the Chinese open source companies together (more ad

industryemad-mostaque--x
22 Jun 2026
Industry

My version of AI mania is when I get a spidey feeling that the models have changed. Opus 4.8 feels very different today.

DGX agent

Allie K. Miller expresses subjective observations about perceiving changes in Claude Opus 4.8's behavior or capabilities, describing an intuitive sense ('spidey feeling') that the model feels notably

industryallie-k--miller--x
22 Jun 2026
Hardware

MaineCoon is the first video model that focuses on social interactions: facial expressions, emotions, fluid conversation, audio-lip sync, et…

DGX agent

MaineCoon is the first video model that focuses on social interactions: facial expressions, emotions, fluid conversation, audio-lip sync, etc. Really impressive inference specs: 22B params, 47.5 FPS o

hardwarefrancois-chollet--x
21 Jun 2026
Industry

API prices of key AI models: US vs China

DGX agent

This post likely compares the pricing of major AI model APIs between the United States and China, highlighting cost differences that reflect different market conditions, regulatory environments, and c

industryclem-delangue--x
19 Jun 2026
Hardware

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly …

DGX agent

BREAKING NEWS: Anthropic's latest model will NOT help you if it thinks your ML research/ML engineering is interesting, and/or will secretly degrade its IQ so that the average engineer won't notice. We

hardwareclem-delangue--x
9 Jun 2026
Tools

@DeepCogito needed sub-500ms time to first token at 1,000+ requests per minute for their frontier reasoning models. Together AI delivered. H…

DGX agent

@DeepCogito needed sub-500ms time to first token at 1,000+ requests per minute for their frontier reasoning models. Together AI delivered. Hear from the Deep Cogito team on what it takes to build fron

toolstogether-ai--x
9 Jun 2026
Hardware

for those keeping track at home it was 34 days between signing this deal and launching Mythos-class model GA to the world. https://x.com/lee…

DGX agent

for those keeping track at home it was 34 days between signing this deal and launching Mythos-class model GA to the world. https://x.com/leerob/status/2052059466821198061?s=20 building on @nvidia stac

hardwareswyx--x
9 Jun 2026
Local Ai

this model is the opposite of mythos. Its small, cost effective, apache 2.0, and locally deployable. This is the way LLMs should go. small, …

DGX agent

this model is the opposite of mythos. Its small, cost effective, apache 2.0, and locally deployable. This is the way LLMs should go. small, open source, transparent and sovereign vs large, expensive,

local-aiclem-delangue--x
9 Jun 2026
Applications

The next quarter of effort at the frontier of applied AI is not about model capabilities. It's about 3 key things: 1. security, guardrails a…

DGX agent

The next quarter of effort at the frontier of applied AI is not about model capabilities. It's about 3 key things: 1. security, guardrails around data ingress/egress, and understanding risk in your sy

applicationslinus-lee--x
8 Jun 2026
Industry

We are releasing our first quantized checkpoints for the Qwen3.5 series of models, co-designed jointly with our inference engine to achieve …

DGX agent

We are releasing our first quantized checkpoints for the Qwen3.5 series of models, co-designed jointly with our inference engine to achieve maximum possible performance on Apple hardware Starting from

industryclem-delangue--x
8 Jun 2026
Tutorials

This is a pretty striking shift toward Chinese models by American AI startups since the start of the year. https://substack.com/@profgmarket…

DGX agent

American AI startups have increasingly shifted toward adopting Chinese AI models since the beginning of the year, representing a notable change in technology sourcing preferences. This trend likely re

tutorialsjeremy-howard--x
7 Jun 2026
Local Ai

Meet LM Studio's mobile app. Your local models, now in your pocket.

DGX agent

LM Studio has released a mobile app that enables users to run and access local language models on their smartphones. The app extends LM Studio's desktop functionality, allowing users to deploy and int

local-ailm-studio--x
4 Jun 2026
Model Releases

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --mo…

DGX agent

ollama run gemma4:12b Gemma 4 12B is updated on Ollama, and available across all platforms! Try it on: Claude Code ollama launch claude --model gemma4:12b Hermes Agent ollama launch hermes --model gem

model-releasesollama--x
4 Jun 2026
Industry

@garrytan Model, application, and infrastructure layers trying to commoditize each other

DGX agent

This post discusses how the model, application, and infrastructure layers of AI are competing to commoditize each other, suggesting a dynamic where each layer is attempting to move upstream or downstr

industrysonya-huang--x
2 Jun 2026
Hardware

Nvidia announcing a 550B model wasn't on my bingo card They are now the strongest american open-source lab

DGX agent

Nvidia announced development of a 550 billion parameter language model, positioning itself as a leading open-source AI research organization competing with traditional academic and independent labs. T

hardwareclem-delangue--x
1 Jun 2026
Applications

I think Epoch does a great job benchmarking, but I continue to believe that open weights models are much more fragile, especially out-of-dis…

DGX agent

I think Epoch does a great job benchmarking, but I continue to believe that open weights models are much more fragile, especially out-of-distribution, than their benchmarks indicate. Vibe-wise, I don’

applicationsethan-mollick--x
30 May 2026
Applications

Different models need different prompts, sometimes tools “Harness profiles” are how we do that in deepagents

DGX agent

Different models need different prompts, sometimes tools “Harness profiles” are how we do that in deepagents Deep Agents v0.6 makes harness profiles a first-class abstraction. Now, you can get product

applicationsharrison-chase--x
29 May 2026
Hardware

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

DGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

hardwareclem-delangue--x
28 May 2026
Research

Would you like to join the research effort on JEPA and World Models easily? After a full year of hard work, we’re excited to finally release…

DGX agent

Would you like to join the research effort on JEPA and World Models easily? After a full year of hard work, we’re excited to finally release stable-worldmodel: an open-source, scalable platform built

researchyann-lecun--x
28 May 2026
Applications

Google has the only true Omni model, but the elements aren't hooked up. It appears it can take in & output audio, images. video, songs, text…

DGX agent

Google has the only true Omni model, but the elements aren't hooked up. It appears it can take in & output audio, images. video, songs, text, code, etc. But right now each type of output is separate.

applicationsethan-mollick--x
27 May 2026
Industry

Great to see @poolsideai (US lab) committing to open sourcing their foundation models going forward Laguna is an interesting release, check …

DGX agent

Great to see @poolsideai (US lab) committing to open sourcing their foundation models going forward Laguna is an interesting release, check it out @Shaughnessy119 https://poolside.ai/blog/introducing-

industryemad-mostaque--x
27 May 2026
Industry

If the Founder of Hugging Face asks, you gotta do it. Models and dataset now live: https://huggingface.co/papers/2605.22391 Also built an ex…

DGX agent

If the Founder of Hugging Face asks, you gotta do it. Models and dataset now live: https://huggingface.co/papers/2605.22391 Also built an explorer: https://huggingface.co/spaces/Kaikaku/epicure-explor

industryclem-delangue--x
27 May 2026
Tutorials

Wow. It looks like the @XiaomiMiMo v2.5 model is insanely good value :O (Price for each prompt shown after each answer. Context includes >40…

DGX agent

Jeremy Howard comments on the Xiaomi MiMo v2.5 model, highlighting its exceptional value proposition and cost-effectiveness for prompt processing. The post appears to include comparative pricing data

tutorialsjeremy-howard--x
27 May 2026
Local Ai

Comfy will be at @aionthelot Our CEO @yoland_yan is joining a panel of AI video leaders for a progress report on the state of video models -…

DGX agent

Comfy will be at @aionthelot Our CEO @yoland_yan is joining a panel of AI video leaders for a progress report on the state of video models - where they excel, where they fall short, and what's coming

local-aicomfyui--x
26 May 2026
Tutorials

// Language Models Need Sleep // Let your agents 'sleep', folks. On a serious note, this is a fascinating paper on getting the most from lon…

DGX agent

// Language Models Need Sleep // Let your agents 'sleep', folks. On a serious note, this is a fascinating paper on getting the most from long-horizon agents. Here is the problem with agents today: Att

tutorialsdair-ai--x
26 May 2026
Local Ai

My daily average local model token burn is 17M They have become tremendously useful.

DGX agent

Clem Delangue reports consuming approximately 17 million tokens daily when running local language models, indicating substantial usage of on-device AI inference. He expresses satisfaction with the uti

local-aiclem-delangue--x
24 May 2026
Research

Thinking Machines is impressive. In a couple hours I just fine tuned my own Qwen3.5-397B model this afternoon. Fast usable multimodal is als…

DGX agent

Thinking Machines is impressive. In a couple hours I just fine tuned my own Qwen3.5-397B model this afternoon. Fast usable multimodal is also going to enable very mind-blowing personal AI. People talk

researchsoumith-chintala--x
24 May 2026
Industry

A 6-person team is building task-specific AI models that are 4-8x faster than anything from OpenAI or Anthropic. 500K downloads on HuggingFa…

DGX agent

A 6-person team is building task-specific AI models that are 4-8x faster than anything from OpenAI or Anthropic. 500K downloads on HuggingFace. No hype. Just better engineering winning on the merits.

industryclem-delangue--x
23 May 2026
← Previous
1…2526272829…128
Next →