Model Releases
GPT-OSS has turned one year old today!
It is one of the best local models ever released, in both 20B and 120B versions. I always come back to it, especially the 120B version. Its only competition is, in my opinion, Qwen 3.5 122B, but that
It is one of the best local models ever released, in both 20B and 120B versions. I always come back to it, especially the 120B version. Its only competition is, in my opinion, Qwen 3.5 122B, but that model is much slower (A10B) and has not been released in a local-friendly QAT format (such as MXFP4). Nemotron 3 Super is disappointing; it is close in capability and feels like a GPT-OSS 120B clone with a better architecture. It is also slower (due to being A12B). NVIDIA essentially created a slower GPT-OSS 120B clone. Mistral 4 Small has similar problems. It is definitely not smarter, although it thinks less (for better or worse). OpenAI made a great model, and I hope they release a successor eventually. In the meantime, you may find my attempt at improving the GPT-OSS Jinja template useful. It is primarily based on the Unsloth version (so tool calls work correctly) and incorporates TypeError fixes from elsewhere, as well as additional sanity checks and configurable token smuggling protection. I hope someone finds my humble contribution useful: https://huggingface.co/arbv/gpt-oss-fixed-jinja-template It's not much, but it's honest work. submitted by /u/arbv [link] [comments]
Related
- Real-world reality check on Qwen for autonomous coding agents
- Now Suddenly too many choices for DGX Spark with Qwen 3.5 122B . What would be the next upgrade?
- [[release-wintermix-qwen35-122b-a10b-in-native-mlx-an-82-gib-b|[Release] WinterMix — Qwen3.5-122B-A10B in native MLX: an 82 GiB build that beats 94–95 GiB quants, plus a 68 GiB build for agent swarms]]
Source: r/LocalLLaMA | 2026-08-04