Model Releases
šØ Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 𤯠Itās an open-source pipelā¦
šØ Are we witnessing the automation of AI research? @HuggingFace just unveiled 'ML-Intern' and my mind is BLOWN 𤯠Itās an open-source pipeline that replicates the exact daily loop of an ML researcher.
šØ Are we witnessing the automation of AI research? @HuggingFace just unveiled "ML-Intern" and my mind is BLOWN 𤯠Itās an open-source pipeline that replicates the exact daily loop of an ML researcher. You simply write a prompt, then watch the magic happen: ā ML-Intern reads the arXiv papers ā digs through citations ā spins up GPU sandboxes ā iterates ā ... even builds you a deeply researched model Awesome, right? Plus, everything happens right inside the HF ecosystem: š¹ training via HF Jobs š¹ monitoring via Trackio š¹ dataset loading & final model pushing straight to the Hub The proof is in the pudding! They let ML-Intern tackle scientific reasoning. That thing autonomously: - researched official benchmarks - found OpenScience and NemoTron-CrossThink - pulled 7 difficulty-filtered variants (ARC/SciQ/MMLU) - ... and ran 12 SFT runs on Qwen3-1.7B The results? GPQA jumped from 10% to 32% in under 10 hours! As a comparison, Claude Code topped out at 22.99%. This secured SOTA on PostTrainBench, the rigorous new benchmark by ELLIS Institute Tübingen and Max Planck Institute for 10-hour LLM post-training. Itās out now, and completely free and open-source. You've got to dig into this code and app below š§µ ā Media
Related
- Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real reseā¦
- Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships ā¦
- HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds reseaā¦
- We need open traces so that everyone can train open agent models! cc @steipete @badlogicgames @thdxr @matanSF @hwchase17
Source: Clem Delangue (X) | 2026-04-22