AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “clem-delangue--x”

GridTimelineEvolution
849 results
17 Apr 2026

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local …

Model ReleasesDGX agent

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local agents are today: # Start llama.cpp server: llama-server -hf

the coding agent traces viewer on @huggingface is very nice! https://huggingface.co/datasets/badlogicgames/pi-mono/blob/main/2026-01-16T11-1…

AgentsDGX agent

The post praises a coding agent tracer/viewer tool available on Hugging Face, highlighting its quality and usefulness for developers working with agents. The tool appears to be part of the pi-mono dat

The king reigns supreme This is likely going to be the best finetune of qwen 3.6 35b https://huggingface.co/DJLougen/Ornstein3.6-35B-A3B-GGU…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

This post announces Ornstein3.6-35B-A3B-GGU, a fine-tuned variant of Qwen 3.6 35B model available on Hugging Face, with the author claiming it represents a high-quality optimization of the base model.

16 Apr 2026

180 tok/s generation on a 4090 with qwen 3.6. if you're on a 4090 and not running this model yet you're leaving performance on the table. 3B…

Model ReleasesDGX agent

180 tok/s generation on a 4090 with qwen 3.6. if you're on a 4090 and not running this model yet you're leaving performance on the table. 3B active params at that speed is insane for agentic coding. t

Adithya is putting his work where his heart is. If you believe in open research and open AI, join a company that actually lives those values…

IndustryDGX agent

Adithya is putting his work where his heart is. If you believe in open research and open AI, join a company that actually lives those values, not a closed-source, revenue-maximizing one! Quick career

Completely agree with Jensen! Not exporting AI out of fear is classic case where the cure would be 100x worse than the disease. You slow dow…

HardwareDGX agent

Completely agree with Jensen! Not exporting AI out of fear is classic case where the cure would be 100x worse than the disease. You slow down innovation, progress and US technology and economic leader

Great blogpost from @pcuenq on making a new skill + test harness to automate porting new models from Transformers to mlx-lm

Local AiDGX agent

This post discusses a blog article by @pcuenq that covers the process of creating new skills and test harnesses to automate the conversion of machine learning models from the Hugging Face Transformers

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Re…

Model ReleasesDGX agent

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to

i feel no shame. zero. none.

AgentsDGX agent

i feel no shame. zero. none. You can now visualize Pi traces that you upload on @huggingface! Let's make sharing agent traces 10x more common to make agent AI more open and collaborative! Also, becaus

Just added 8 new developer-loved providers to @stripe projects: > stripe projects catalog > @huggingface, @Cloudflare, @OpenRouter, @firecra…

IndustryDGX agent

Just added 8 new developer-loved providers to @stripe projects: > stripe projects catalog > @huggingface, @Cloudflare, @OpenRouter, @firecrawl, @flydotio, @Amplitude_HQ, @mixpanel, @inngest Your agent

Just launched some great new providers in Stripe Projects (http://projects.dev): @huggingface @Cloudflare @OpenRouter @firecrawl @flydotio @…

IndustryDGX agent

Just launched some great new providers in Stripe Projects (http://projects.dev): @huggingface @Cloudflare @OpenRouter @firecrawl @flydotio @Amplitude_HQ @mixpanel @inngest Many more coming later this

LeRobot is the unsung hero of open robotics!

IndustryDGX agent

LeRobot is an open-source robotics framework or platform that aims to democratize access to robotics development and research. The post suggests it addresses gaps in the robotics field by providing ac

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than bas…

Model ReleasesDGX agent

New insane model from Jackrong on @huggingface 🤯 Qwen3.5-9B-GLM5.1-Distill-v1 🧠 Distilled on GLM-5.1 reasoning ⚙️ Deeper thinking than base model 🧪 Benchmarks coming soon ✅ Fits on 8GB VRAM ✍️ New mod

Open source software will be many times more secure than closed source software in the new Mythos era

IndustryDGX agent

Open source software will be many times more secure than closed source software in the new Mythos era they: OpenClaw is so insecure look at all these GHSAs! reality: we are just an indicator of the co

انتهت حجة 'ما عندي كرت شاشة قوي عشان أشغل ذكاء اصطناعي محلي'. علي بابا نزلت موديل Qwen3.6 بهندسة MoE. حجمه الكلي 35B (يعطيك ذكاء عالي جداً)،…

AgentsDGX agent

انتهت حجة 'ما عندي كرت شاشة قوي عشان أشغل ذكاء اصطناعي محلي'. علي بابا نزلت موديل Qwen3.6 بهندسة MoE. حجمه الكلي 35B (يعطيك ذكاء عالي جداً)، لكن وقت الـ Inference يستهلك 3B فقط! يعني أداء Enterprise ع

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I d…

Model ReleasesDGX agent

Shocking result on my pelican benchmark this morning, I got a better pelican from a 21GB local Qwen3.6-35B-A3B running on my laptop than I did from the new Opus 4.7! Qwen on the left, Opus on the righ

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing!

Model ReleasesDGX agent

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing! ⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agen

The open-source AI community just got a new home for their data workflows. 🤗 @huggingface is now available in Adaptive Data. Pull datasets …

IndustryDGX agent

The open-source AI community just got a new home for their data workflows. 🤗 @huggingface is now available in Adaptive Data. Pull datasets directly into a platform that evolves with the problems you'r

this is the dataset btw: https://huggingface.co/datasets/badlogicgames/pi-mono

IndustryDGX agent

The pi-mono dataset, hosted on Hugging Face by badlogicgames, is a machine learning dataset likely containing monophonic audio or single-channel data related to pi (π) mathematical constants or Pi-rel

WAIT WHAT?! 2-bit Qwen3.6-35B-A3B is lightning fast and it only needs 13 GB RAM. “did a complete repo bug hunt with evidence, repro, fixes, …

IndustryDGX agent

A developer successfully optimized Qwen 3.6-35B model to 2-bit quantization, achieving significant performance improvements with only 13GB RAM requirements while maintaining functionality. The work in

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to valida…

Model ReleasesDGX agent

We replicated Mythos findings in opencode using public models, not Anthropic's private stack. The moat is moving from model access to validation: finding vulnerability signal is getting cheaper; turni

We’re open-sourcing HY-World 2.0, a multimodal world model that generates, reconstructs, and simulates interactive *3D worlds* from text, im…

ApplicationsDGX agent

We’re open-sourcing HY-World 2.0, a multimodal world model that generates, reconstructs, and simulates interactive *3D worlds* from text, images, and videos. Outputs can be integrated into game engine

You can now visualize Pi traces that you upload on @huggingface! Let's make sharing agent traces 10x more common to make agent AI more open …

AgentsDGX agent

You can now visualize Pi traces that you upload on @huggingface! Let's make sharing agent traces 10x more common to make agent AI more open and collaborative! Also, because it's fun to analyze @badlog

15 Apr 2026

Given the increasingly closed-source nature of the U.S. AI ecosystem, it is now more important than ever to push for the proliferation of op…

IndustryDGX agent

Given the increasingly closed-source nature of the U.S. AI ecosystem, it is now more important than ever to push for the proliferation of open model and dataset releases. Datamule (@johngfriedman), @T

More people should work on harnesses for open and local models!

IndustryDGX agent

Hugging Face co-founder and CEO Clément Delangue advocates for more developers and researchers to focus on building evaluation harnesses and testing frameworks specifically designed for open-source an

Open-source is the solution to cyber-security because with new AI capabilities, all of the open-source repos will be inspected and patched 1…

AgentsDGX agent

Open-source is the solution to cyber-security because with new AI capabilities, all of the open-source repos will be inspected and patched 100x faster/better than any closed-source system! Let's take

SONIC training code + Finetuning checkpoint + VLA data collection scripts are open-sourced. Little easter egg on the GEAR-SONIC website too …

IndustryDGX agent

SONIC training code + Finetuning checkpoint + VLA data collection scripts are open-sourced. Little easter egg on the GEAR-SONIC website too :) https://github.com/NVlabs/GR00T-WholeBodyControl SONIC is

The most powerful LLM to run at home:

IndustryDGX agent

The most powerful LLM to run at home: 소형 로컬LLM 중 가장 강력한 모델을 소개합니다. 🔥SuperGemma4-31b-abliterated 우리가 원하는 로컬 모델의 모든것 - 무검열, 가벼움, 똑똑함 벤치마크 평가 : MMLU, GPQA, IFEval 등 벤치 항목을 종합하여 판단. 모델의 태생 약점, 비효율적인 연산단계와

14 Apr 2026

Boom, the game is changed GLM 5.1 running locally seems to actually work… Now I can run a lot of my @openclaw workflows for the cost of elec…

IndustryDGX agent

Clem Delangue (or someone in their network) shared excitement about successfully running GLM 5.1 locally, noting it appears to be functional and effective. The post highlights the cost advantage of ru

I just updated our license. For personal use, you’re free to run the software on your own servers for coding, building applications, agents,…

IndustryDGX agent

I just updated our license. For personal use, you’re free to run the software on your own servers for coding, building applications, agents, tools, or integrations, as well as for research, experiment

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GP…

HardwareDGX agent

Introducing Kernels on the Hugging Face Hub ✨ What if shipping a GPU kernel was as easy as pushing a model? - Pre-compiled for your exact GPU, PyTorch & OS - Multiple kernel versions coexist in one pr

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, th…

AgentsDGX agent

Is there somewhere a collection of the best agent/coding harnesses for each models, especially open-source and local ones? In my opinion, the biggest reason why people are struggling with open/local m

just getting started

IndustryDGX agent

This page indicates that the functionality of x.com is restricted due to technical issues on the user's end. Primary issues include JavaScript being disabled in the browser, requiring the user to enab

let's go open-source and local models!

Model ReleasesDGX agent

let's go open-source and local models! Uber's CTO told @LauraBratton5 that AI coding tools—particularly Anthropic’s Claude Code—has already maxed out its 2026 AI budget 📈 “I'm back to the drawing boar

llama.cpp built from source + qwen3.5 27b running locally + hermes agent on top + camofox for scraping (tls spoofing) + scweet to scrape x (…

Model ReleasesDGX agent

llama.cpp built from source + qwen3.5 27b running locally + hermes agent on top + camofox for scraping (tls spoofing) + scweet to scrape x (no api keys) + tailscale to access from other devices bro Me

M2.7 w/ hermes cli is replacing ~75% of my claude code / opus usage now, but we need clarity for using it as a coding agent @ work. We're tr…

Model ReleasesDGX agent

M2.7 w/ hermes cli is replacing ~75% of my claude code / opus usage now, but we need clarity for using it as a coding agent @ work. We're truly blessed to have the weights of this one, looking forward

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call na…

IndustryDGX agent

New post: We show that small, cheap models can detect the flagship Mythos FreeBSD zero-day (CVE-2026-4747) using a simple harness we call nano-analyzer Models down to 3.6B active params (including ope

Randomly made the HN frontpage with my latest blog post on function calling and open source models. https://www.thetypicalset.com/blog/gramm…

IndustryDGX agent

Remi Louf shared that a blog post he wrote about function calling and open source models unexpectedly reached the Hacker News frontpage. The post, hosted on thetypicalset.com, likely explores how open

Sub-32B open weights models now offer GPT-5 level intelligence with Qwen3.5 27B (Reasoning) matching GPT-5 (medium) at 42 and Gemma 4 31B (R…

Model ReleasesDGX agent

Sub-32B open weights models now offer GPT-5 level intelligence with Qwen3.5 27B (Reasoning) matching GPT-5 (medium) at 42 and Gemma 4 31B (Reasoning) matching GPT-5 (low) at 39 on the Artificial Analy

Time to follow http://hf.co/tencent!

IndustryDGX agent

Time to follow http://hf.co/tencent! Genie3 generates videos. We generate 𝟯𝗗 𝘄𝗼𝗿𝗹𝗱𝘀 you can actually use. Launching tomorrow — Tencent #HYWorld 2.0, an engine-ready World Model🚀 This isn't a video. It

13 Apr 2026

Benchmarked @DJLougen ’s Ornstein-27B-v2 Q6_K on my RTX 3090 using hermes-bench, my new open-source benchmarking UI for local LLMs and Herme…

Model ReleasesDGX agent

Benchmarked @DJLougen ’s Ornstein-27B-v2 Q6_K on my RTX 3090 using hermes-bench, my new open-source benchmarking UI for local LLMs and Hermes agents. Ornstein is a Qwen 3.5 27B fine-tune trained on re

DFlash for Kimi-K2.5 was pushed 3 hours ago! Acceptance length varies from 4.0 to 6.3 on datasets. This is only with SGLang btw. https://hug…

IndustryDGX agent

DFlash, an optimized inference technique, has been integrated for the Kimi-K2.5 model and pushed to SGLang approximately 3 hours prior to the post. The implementation achieves acceptance lengths rangi

🚨 SUPER GEMMA 4 26B UNCENSORED IS INSANE LLM WIZARD COOKING AGAIN @songjunkr Dropped SuperGemma4-26B-Uncensored GGUF v2 and it’s trending o…

Model ReleasesDGX agent

🚨 SUPER GEMMA 4 26B UNCENSORED IS INSANE LLM WIZARD COOKING AGAIN @songjunkr Dropped SuperGemma4-26B-Uncensored GGUF v2 and it’s trending on @huggingface🤗 This thing SMOKES the regular Gemma-4 26B: 🤯0

Trending well! We're glad the traces are useful to the @NousResearch Hermes Agent community. A third batch is in progress.

Model ReleasesDGX agent

Trending well! We're glad the traces are useful to the @NousResearch Hermes Agent community. A third batch is in progress. Very cool open-source traces from @TheZachMueller @LambdaAPI: https://hugging

We just OCR'd 27,000 arxiv papers into Markdown using an open 5B model, 16 parallel HF Jobs on L40S GPUs, and a mounted bucket. Total cost: …

IndustryDGX agent

We just OCR'd 27,000 arxiv papers into Markdown using an open 5B model, 16 parallel HF Jobs on L40S GPUs, and a mounted bucket. Total cost: $850 Total time: ~29 hours Jobs that crashed: 0 This now pow

we used Chandra-OCR-2 by @datalabto: https://huggingface.co/datalab-to/chandra-ocr-2 Full write-up by @NielsRogge: https://huggingface.co/bl…

IndustryDGX agent

Chandra-OCR-2 is an optical character recognition model developed by DataLab, available on Hugging Face at datalab-to/chandra-ocr-2. The model was highlighted by Hugging Face CEO Clément Delangue, wit

We've partnered with @huggingface and @pollenrobotics to bring Gemini Live to the Reachy Mini conversation app 🤖💬 Thanks so much @ailozovs…

Model ReleasesDGX agent

Google's Gemini Live has been integrated into the Reachy Mini robot's conversation application through a partnership between Hugging Face, Pollen Robotics, and Google. This collaboration brings advanc

12 Apr 2026

We're delighted to announce that MiniMax M2.7 is now officially open source. With SOTA performance in SWE-Pro (56.22%) and Terminal Bench 2 …

IndustryDGX agent

We're delighted to announce that MiniMax M2.7 is now officially open source. With SOTA performance in SWE-Pro (56.22%) and Terminal Bench 2 (57.0%). You can find it on Hugging Face now. Enjoy!🤗 huggin

10 Apr 2026

favorite AGI/sci-fi vibe these days is coding a robot code together with the robot here vibe-pluging @ElevenLabs in @reachymini for a talk l…

IndustryDGX agent

Thomas Wolf, co-founder and Chief Science Officer of Hugging Face, shared a post describing his experience 'vibe coding' robot behaviors collaboratively with a Reachy Mini robot — an open-source d...

hermes agent traces on Hf would be pretty neat attn: @Teknium @NousResearch @ClementDelangue https://huggingface.co/changelog/agent-trace-vi…

AgentsDGX agent

Hugging Face now supports uploading agent traces directly to HF Datasets, where the Hub auto-detects trace formats and tags datasets as 'Traces,' providing a dedicated viewer for browsing sessions...

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…

Model ReleasesDGX agent

In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h

Ran autoresearch on hf to see whether anything can beat MuonAdamW baseline Biggest takeaway: NS orthogonalization is a very strong attractor…

IndustryDGX agent

Ran autoresearch on hf to see whether anything can beat MuonAdamW baseline Biggest takeaway: NS orthogonalization is a very strong attractor that absorbs most gradient modifications you throw at it. S

Should we start an open Glasswing?

IndustryDGX agent

The specific X (Twitter) post by Clément Delangue asking 'Should we start an open Glasswing?' could not be directly accessed, and search results did not surface the precise content or context of th...

The reaction people are having to AIs that can find bugs in code is fascinating. Finally, we have the capacity to fix the crisis in computer…

SafetyDGX agent

The reaction people are having to AIs that can find bugs in code is fascinating. Finally, we have the capacity to fix the crisis in computer security we’ve had for decades, and everyone is treating it

9 Apr 2026

Did the famous @huggingface HQ tour in Paris! Huge thanks to @xenovacom 😊 Such an amazing space!

IndustryDGX agent

The specific tweet (status/2042207101557342221) is not directly accessible or indexed in search results, and the URL itself appears to be from a future date beyond current indexing. However, based ...

Do you understand what this means? For the first time, an Open Weight models is #1 on CyberSecutity. Sure there’s Mythos but we don’t have i…

IndustryDGX agent

The specific tweet from @0xSero is not directly accessible, and the search results don't surface the exact model or event being referenced in that post. However, based on the context clues in the t...

Feels like one of the cybersecurity risks over the coming months will be widely used open-source projects that are simply too lightly mainta…

IndustryDGX agent

Feels like one of the cybersecurity risks over the coming months will be widely used open-source projects that are simply too lightly maintained for how critical they’ve become. A few ways to help: -

Having a model like Gemma 4, which is perfectly adequate for everyday use in many cases, runs locally, is free, and secure, still feels unre…

Model ReleasesDGX agent

Having a model like Gemma 4, which is perfectly adequate for everyday use in many cases, runs locally, is free, and secure, still feels unreal. We have a very good AI that costs nothing, uses hardly a

Someone used Pi to build an Excel agent… then published all the traces from it to HuggingFace🤯 Just point your agent at this changelog and …

AgentsDGX agent

Someone used Pi to build an Excel agent… then published all the traces from it to HuggingFace🤯 Just point your agent at this changelog and tell it to publish its traces to HuggingFace! http://huggingf

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at…

HardwareDGX agent

This is the full video of the hardest version of the task: t-shirt folding from unstructured initial states. This setting really requires at least some strategy, since the robot first has to spread th

← Previous
1…12131415
Next →