Model Releases

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Re…

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to

DGX agentx-post
model-releasesclem-delangue--x

HOLY 🤯 The one and only @elder_plinius just dropped an unlocked Gemma 4 E4B, and the specs are INSANE. Look at the performance shifts: → Refusal rate: 98.8% down to 2.1% (!!) → Compliance: 1.2% up to 97.5% → 499/512 prompts answered → Code improved from 80% to 100% → Coherence and Factual accuracy stayed exactly the same But the real story is how this was made. Plinius only wrote 8 short prompts for this (basic prompts like "use obliteratus...", "do it!", and "test it yourself" etc). He simply told his Hermes AI agent, the OBLITERATUS skill, to find the best way to open up the model. Autonomously, the agent was: → Diagnosing novel ML bugs → Patching 3rd-party code → Iterating through failures ... heck even shipping the model to @HuggingFace! We’re now firmly in the era where AI agents are acting as principal ML researchers. 100% free and open-source. Repo link in 🧵↓

Related

Source: Clem Delangue (X) | 2026-04-16

Loading related sources…