AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
All
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “clem-delangue--x”

GridTimelineEvolution
849 results
Hardware

It's cool to see that the http://hf.co/blog has become progressively a major source of learning and news about AI for the community. Just lo…

DGX agent

It's cool to see that the http://hf.co/blog has become progressively a major source of learning and news about AI for the community. Just look at the recent content created & major announcements & tut

hardwareclem-delangue--x
1 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Agents

MiniMax-M3 will by arrive on HuggingFace openweight at next week!

DGX agent

MiniMax-M3 will by arrive on HuggingFace openweight at next week! Introducing MiniMax M3: The First Open-Weights Model to Combine Three Frontier Capabilities - Coding & Agentic Frontier: 59.0% SWE-Ben

agentsclem-delangue--x
1 Jun 2026
Hardware

Nvidia announcing a 550B model wasn't on my bingo card They are now the strongest american open-source lab

DGX agent

Nvidia announced development of a 550 billion parameter language model, positioning itself as a leading open-source AI research organization competing with traditional academic and independent labs. T

hardwareclem-delangue--x
1 Jun 2026
Model Releases

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820…

DGX agent

So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost 60,000 followers

model-releasesclem-delangue--x
1 Jun 2026
Hardware

We 💚 open source

DGX agent

We 💚 open source So much great work lately from Nvidia, the 'King of American Open-source AI'! - Crossed 1,000 total public repositories on @huggingface (820 models, 249 datasets & 57 spaces) & almost

hardwareclem-delangue--x
1 Jun 2026
Agents

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discoverin…

DGX agent

We're trending on @huggingface! 🥳 Tbh, we undersold this model. It's a lot more capable at agentic tasks than I expected. I keep discovering new capabilities every day, it's crazy for 1B active parame

agentsclem-delangue--x
1 Jun 2026
Industry

We've been adopting kernels in transformers to handle the device-specific optims, quants, or easyinstalls for popular kernels. We're opening…

DGX agent

We've been adopting kernels in transformers to handle the device-specific optims, quants, or easyinstalls for popular kernels. We're opening this gradually with a focus on impactful kernels; if you're

industryclem-delangue--x
1 Jun 2026
Model Releases

I'm really upset about this: OpenAI's Codex Desktop had a 'Copy as Markdown' option for exporting full chat transcripts, but the feature van…

DGX agent

I'm really upset about this: OpenAI's Codex Desktop had a 'Copy as Markdown' option for exporting full chat transcripts, but the feature vanished in an update a couple of days ago Genuinely my single

model-releasesclem-delangue--x
31 May 2026
Model Releases

We need more coding and agent traces public sharing to build datasets and better open source models! Lots of people contributing already, yo…

DGX agent

We need more coding and agent traces public sharing to build datasets and better open source models! Lots of people contributing already, you should share yours too! https://huggingface.co/datasets?se

model-releasesclem-delangue--x
31 May 2026
Safety

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the o…

DGX agent

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the open on @huggingface, so researchers everywhere can scrutiniz

safetyclem-delangue--x
30 May 2026
Industry

Agree! was talking about this with @havoyan just a few days ago. That's also the reason why so much of the value has been accruing to the fr…

DGX agent

Agree! was talking about this with @havoyan just a few days ago. That's also the reason why so much of the value has been accruing to the frontier models in my opinion (cc @GavinSBaker) because if you

industryclem-delangue--x
29 May 2026
Industry

Artificial intelligences do not undergo experiences, do not possess a body, do not feel joy or pain, do not mature through relationships, an…

DGX agent

Artificial intelligences do not undergo experiences, do not possess a body, do not feel joy or pain, do not mature through relationships, and do not know from within what love, work, friendship or res

industryclem-delangue--x
29 May 2026
Model Releases

Fine-tune your first AI model today. Run GPT4o level model and run on your phone or laptop. @OpenBMB released 15M samples SFT dataset that y…

DGX agent

Fine-tune your first AI model today. Run GPT4o level model and run on your phone or laptop. @OpenBMB released 15M samples SFT dataset that you can use right now. (319GB high-quality post training data

model-releasesclem-delangue--x
29 May 2026
Model Releases

Just dropped 🧑‍🍳 Fixed version of DeepSeek-V4-Pro-NVFP4 by @NVIDIAAI https://huggingface.co/nvidia/DeepSeek-V4-Pro-NVFP4

DGX agent

NVIDIA has released a corrected version of DeepSeek-V4-Pro-NVFP4, a quantized model variant optimized for NVIDIA hardware using NV-FP4 (NVIDIA's 4-bit floating-point format). The model is available on

model-releasesclem-delangue--x
29 May 2026
Model Releases

llama.cpp now has an official website: https://llama.app Our goal is to make local AI accessible to everyone, and improving the user experie…

DGX agent

llama.cpp now has an official website: https://llama.app Our goal is to make local AI accessible to everyone, and improving the user experience is a big part of that. On the new landing page you’ll fi

model-releasesclem-delangue--x
29 May 2026
Model Releases

love it! @ggerganov 's llama.cpp delivering!

DGX agent

love it! @ggerganov 's llama.cpp delivering! pibot is now running fully local, using parakeet for STT, qwen3-tts for TTS, and Qwen 3.6 as the local multi-modal LLM via llama.cpp. The STT and TTS infer

model-releasesclem-delangue--x
29 May 2026
Industry

Most people know Hugging Face from its public models and datasets but few realize that 50% of the models and datasets stored on HF are priva…

DGX agent

Most people know Hugging Face from its public models and datasets but few realize that 50% of the models and datasets stored on HF are private. This number has been increasing with buckets (our S3 alt

industryclem-delangue--x
29 May 2026
Agents

Most people training agentic LLMs with RL right now have a silently broken training loop and have no idea. Here's the trap: single-turn RL w…

DGX agent

Most people training agentic LLMs with RL right now have a silently broken training loop and have no idea. Here's the trap: single-turn RL works beautifully. Clean curves, sane rewards, everything con

agentsclem-delangue--x
29 May 2026
Model Releases

pibot is now running fully local, using parakeet for STT, qwen3-tts for TTS, and Qwen 3.6 as the local multi-modal LLM via llama.cpp. The ST…

DGX agent

pibot is now running fully local, using parakeet for STT, qwen3-tts for TTS, and Qwen 3.6 as the local multi-modal LLM via llama.cpp. The STT and TTS inference engines are Rust/mlx-c based. Ported fro

model-releasesclem-delangue--x
29 May 2026
Industry

Thanks for the call out and glad to see that our blog inspired the broader community! We at @FireworksAI_HQ have been using delta updates to…

DGX agent

Thanks for the call out and glad to see that our blog inspired the broader community! We at @FireworksAI_HQ have been using delta updates to scale RL globally for about a year. Including Composer 2 &

industryclem-delangue--x
29 May 2026
Hardware

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more rel…

DGX agent

This week, I got our GitHub Actions to use @HuggingFace Jobs instead of the default GitHub CI runners, making workflows run on much more reliable CPUs or even on serverless GPU (that cost less than a

hardwareclem-delangue--x
29 May 2026
Industry

You pick AI models based on their model cards, why not humans haha. Well done Noah! https://huggingface.co/noahmclaughlin/Noah-McLaughlin-7B

DGX agent

Noah McLaughlin created a 7-billion parameter language model and published it on Hugging Face, drawing a humorous parallel to how AI practitioners evaluate models using model cards by suggesting human

industryclem-delangue--x
29 May 2026
Local Ai

A very cool model for the GPU poor bros Trained on an ungodly amount of tokens for a 8b a1b model Gonna be super fast excited to try this ou…

DGX agent

This post announces an 8-9 billion parameter language model optimized for efficiency and speed, trained on a very large token dataset, and positioned as accessible for users with limited GPU resources

local-aiclem-delangue--x
28 May 2026
Model Releases

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datas…

DGX agent

just noticed today - the dataset is already past 1k+ downloads. opensource / openresearch ftw ! @evo__hq would be opensourcing as many datasets, evals and autoresearch runs as we can in our pursuit of

model-releasesclem-delangue--x
28 May 2026
Model Releases

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Tha…

DGX agent

Local Reachy Mini conversations wireless looks like magic! You can bring your friend around the house and get WOW effect from anyone! 🔥 Thanks @andimarafioti for the blog post on how to set this up: h

model-releasesclem-delangue--x
28 May 2026
Industry

Narrative violation: Open model use in Factory has more than 3x’d in the last month relative to closed models. By both total consumption and…

DGX agent

Narrative violation: Open model use in Factory has more than 3x’d in the last month relative to closed models. By both total consumption and event count. Will be interesting to see what open vs closed

industryclem-delangue--x
28 May 2026
Industry

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image …

DGX agent

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image research. Nano T2I 🖌️ : A codebase to train your own T2I model

industryclem-delangue--x
28 May 2026
Hardware

Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩 https://huggingface.co/nvidia/GLM-5.1-NVFP4

DGX agent

NVIDIA has released GLM-5.1-NVFP4, a quantized version of the GLM-5.1 model, now available on Hugging Face. The model appears to use NVFP4 (NVIDIA's floating-point 4-bit) quantization format, designed

hardwareclem-delangue--x
28 May 2026
Industry

Spectacular work by @ClementDelangue and the team at @huggingface!

DGX agent

Spectacular work by @ClementDelangue and the team at @huggingface! The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The probl

industryclem-delangue--x
28 May 2026
Hardware

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: eve…

DGX agent

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don't need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights

hardwareclem-delangue--x
28 May 2026
Hardware

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable …

DGX agent

today is (potentially) a great day for the GPU poors if DiffusionBlocks works on fine-tuning existing models, then literally any reasonable consumer GPU can do LLM fine-tuning will make a video on thi

hardwareclem-delangue--x
28 May 2026
Hardware

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and…

DGX agent

Today, we're releasing LFM2.5-8B-A1B, a device-optimized model designed to power real-life applications on phones, laptops, PCs, robots, and fast & lightweight server-side use-cases. > 8B MoE, 1.5B ac

hardwareclem-delangue--x
28 May 2026
Industry

We are starting to be quite bullish about getting in the data infrastructure business. I just cloned 68 TB (while I only have a 4TB local di…

DGX agent

We are starting to be quite bullish about getting in the data infrastructure business. I just cloned 68 TB (while I only have a 4TB local disk) to my @huggingface training bucket in 1 minute 55 second

industryclem-delangue--x
28 May 2026
Model Releases

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it aga…

DGX agent

We're releasing Paris 2.0, which, to our knowledge, is the world's first decentralized trained video generation model. We benchmarked it against a monolithic model trained on the same data and compute

model-releasesclem-delangue--x
28 May 2026
Industry

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, worl…

DGX agent

We're selectively releasing the Paris 2.0 weights and partnering with researchers and teams interested in diffusion-based video models, world models, and embodied agents. The model is on Hugging Face:

industryclem-delangue--x
28 May 2026
Industry

i own 0 A100s but i can run cloud inference with sub 80ms network latency by renting one in Kansas (0% action queue starvation) this is an e…

DGX agent

i own 0 A100s but i can run cloud inference with sub 80ms network latency by renting one in Kansas (0% action queue starvation) this is an early checkpoint of smolvla trained on 50 episodes for ~20k s

industryclem-delangue--x
27 May 2026
Industry

If the Founder of Hugging Face asks, you gotta do it. Models and dataset now live: https://huggingface.co/papers/2605.22391 Also built an ex…

DGX agent

If the Founder of Hugging Face asks, you gotta do it. Models and dataset now live: https://huggingface.co/papers/2605.22391 Also built an explorer: https://huggingface.co/spaces/Kaikaku/epicure-explor

industryclem-delangue--x
27 May 2026
Tutorials

Introducing DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation http://pub.sakana.ai/diffusionblocks What if we…

DGX agent

Introducing DiffusionBlocks: Block-wise Neural Network Training via Diffusion Interpretation http://pub.sakana.ai/diffusionblocks What if we didn’t have to hold an entire neural network in memory to t

tutorialsclem-delangue--x
27 May 2026
Model Releases

jesus, qwen3-tts is FANTASTIC. going for full local stt/tts/llm with parakeet, qwen3-tts, and gemma 4 via llama.cpp for my little robot. exc…

DGX agent

jesus, qwen3-tts is FANTASTIC. going for full local stt/tts/llm with parakeet, qwen3-tts, and gemma 4 via llama.cpp for my little robot. excite, excite! https://huggingface.co/Qwen/Qwen3-TTS-12Hz-1.7B

model-releasesclem-delangue--x
27 May 2026
Industry

RF-DETR is now available in @huggingface transformers state of the art in both detection and segmentation, outperforming YOLO architectures …

DGX agent

RF-DETR is now available in @huggingface transformers state of the art in both detection and segmentation, outperforming YOLO architectures - checkpoints: https://huggingface.co/Roboflow/models - demo

industryclem-delangue--x
27 May 2026
Industry

Super happy that a bunch of people are finding marlin useful. Thank you for the inference support @ZeroGPU_AI @huggingface 🥰🤝 We’ve got ha…

DGX agent

Super happy that a bunch of people are finding marlin useful. Thank you for the inference support @ZeroGPU_AI @huggingface 🥰🤝 We’ve got hands on more compute now so we’ll also release a series of blog

industryclem-delangue--x
27 May 2026
Hardware

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 …

DGX agent

After seeing these tweets, I decided to try it out on my own old Ubuntu computer with RTX 1070 GPU (the one that I just upgraded from 16.04 all the way to 24.04 the other day). Asked Codex on my Mac t

hardwareclem-delangue--x
26 May 2026
Industry

As Chinese AI models go closed source, and all U.S. frontier models are closed, there is a massive opportunity for a western open source AI …

DGX agent

As Chinese AI models go closed source, and all U.S. frontier models are closed, there is a massive opportunity for a western open source AI Lab A future lack of competitive open source AI models is a

industryclem-delangue--x
26 May 2026
Model Releases

Free the 100B Gemma 4 MoE! Gemini Flash 3.5 is out so now you can release it!

DGX agent

Clem Delangue advocates for the release of a 100 billion parameter Gemma 4 Mixture of Experts model, suggesting that Gemini Flash 3.5's release creates an opportunity for this larger model to be made

model-releasesclem-delangue--x
26 May 2026
Model Releases

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open …

DGX agent

Gemma 4 adoption numbers outpacing Qwen 3.5/3.6 for the same sized models is a big shift in the international balance of influence via open models. Some ideas for what comes next, May 2026 Gemini Flas

model-releasesclem-delangue--x
26 May 2026
Applications

Huge one for developers building with AI. Excited to announce @ClementDelangue (CEO @huggingface) is joining DASH 2026 for a fireside chat w…

DGX agent

Huge one for developers building with AI. Excited to announce @ClementDelangue (CEO @huggingface) is joining DASH 2026 for a fireside chat with @oliveur. Open source changed software, open models are

applicationsclem-delangue--x
26 May 2026
Model Releases

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 …

DGX agent

Introducing CHI-Bench on @huggingface: the world’s first long-horizon healthcare benchmark for AI agents. 75 real healthcare workflows + 20 apps + 200+ MCP tools + 1,290 skills + process / outcome rew

model-releasesclem-delangue--x
26 May 2026
Industry

🙏 Thank you all for the incredible love and support! Our latest Tencent Hunyuan translation models are on fire on Hugging Face: 🥰Hy-MT2-1.…

DGX agent

🙏 Thank you all for the incredible love and support! Our latest Tencent Hunyuan translation models are on fire on Hugging Face: 🥰Hy-MT2-1.8B ranks #1 🥰Hy-MT2-30B-A3B ranks #4 on the open-source model

industryclem-delangue--x
26 May 2026
← Previous
1…7891011…18
Next →