Local Ai
Call the homies, new @UnslothAI NVFP4 just dropped 🔥
Call the homies, new @UnslothAI NVFP4 just dropped 🔥 We’re releasing new Qwen3.6 quants that run 2.5× faster on your GPU. Qwen3.6-27B NVFP4 runs on 24GB VRAM. 35B-A3B can hit 17,561 tok/s (B200). We a
Call the homies, new @UnslothAI NVFP4 just dropped 🔥 We’re releasing new Qwen3.6 quants that run 2.5× faster on your GPU. Qwen3.6-27B NVFP4 runs on 24GB VRAM. 35B-A3B can hit 17,561 tok/s (B200). We also improved accuracy, tool calling, agent use, and looping. Guide: https://unsloth.ai/docs/models/qwen3.6#nvfp4 Qwen3.6 NVFP4: https…
Source: Clem Delangue (X) | 2026-07-10