Model Releases

GPT-OSS has turned one year old today!

It is one of the best local models ever released, in both 20B and 120B versions. I always come back to it, especially the 120B version. Its only competition is, in my opinion, Qwen 3.5 122B, but that

DGX agentreddit
model-releasesr-localllama

It is one of the best local models ever released, in both 20B and 120B versions. I always come back to it, especially the 120B version. Its only competition is, in my opinion, Qwen 3.5 122B, but that model is much slower (A10B) and has not been released in a local-friendly QAT format (such as MXFP4). Nemotron 3 Super is disappointing; it is close in capability and feels like a GPT-OSS 120B clone with a better architecture. It is also slower (due to being A12B). NVIDIA essentially created a slower GPT-OSS 120B clone. Mistral 4 Small has similar problems. It is definitely not smarter, although it thinks less (for better or worse). OpenAI made a great model, and I hope they release a successor eventually. In the meantime, you may find my attempt at improving the GPT-OSS Jinja template useful. It is primarily based on the Unsloth version (so tool calls work correctly) and incorporates TypeError fixes from elsewhere, as well as additional sanity checks and configurable token smuggling protection. I hope someone finds my humble contribution useful: https://huggingface.co/arbv/gpt-oss-fixed-jinja-template It's not much, but it's honest work. submitted by /u/arbv [link] [comments]

Related

Source: r/LocalLLaMA | 2026-08-04

Loading related sources…