Model Releases
In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for yo…
In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be h
In a world where writing code to build websites and apps is trivial (thank you Lovable, Cursor, Claude,...), the real differentiation for you and your company (and what makes you successful) will be how you manage to train, run and optimize AI models yourself. That's why at Hugging Face, we're doubling down on enabling more to become AI builders rather than AI users. We're releasing this week Kernels on the Hugging Face hub. This repo type is for the hardcore AI engineers among you. Kernels are collections of optimized binary operations where hardware providers support is a first-class citizen: - CUDA - ROCm - Apple Silicon - Intel XPU Expect to see more of this repo type on Hugging Face in the coming days. Featured here: the Flash Attention kernel from @sgl_project team ❤️
Related
- Evaluating LLM-Based 0-to-1 Software Generation in End-to-End CLI Tool Scenarios
- I spent the night testing open-source coding models against Claude Opus in production. Same codebase. Same tasks. Real API calls, real file …
- what he said 🗣️ the very best agents today obsessively tailor the harness layer around the model I’m looking at you “5 things I learned fro…
- 'Claude Code isn't magic. The harness layer is just software, and software is something any dev can shape to fit how they want to work.' Che…
- If you want an example of what this looks like in practice, check out the '/research-docs' skill I created for Claude Code https://x.com/jer…
Source: Clem Delangue (X) | 2026-04-10