Model Releases

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite d…

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite despite requiring 12GB VRAM instead of 24GB 🤯 Runs on a SINGL

DGX agentx-post
model-releasesclem-delangue--x

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite despite requiring 12GB VRAM instead of 24GB 🤯 Runs on a SINGLE RTX 3060 (!) 🧠 Opus 4.6 & GLM-5.1 reasoning in one model ⚡️ 66+ tok/s stable on mid-range GPUs 🧪 Experimental, no additional training 🛠️ Perfect tool calling & agentic reasoning 📷 Fits on low hardware, any 12gb card 📚 GGUF size is 9.8GB (Q4_K_M) Another gift from Jackrong, adding both qwopus and glm-distilled qwen together was not on my bingo card. Truly seems like the sweet spot between 9B and 27B models right now. The ultimate model for 12-16GB VRAM owners? https://huggingface.co/Jackrong/Qwopus-GLM-18B-Merged-GGUF

Source: Clem Delangue (X) | 2026-04-18

Loading related sources…