Industry
New Dflash drafting model for the 27b Lets gooooo https://huggingface.co/z-lab/Qwen3.6-27B-DFlash
A new Dflash drafting model based on Qwen 3.6 with 27 billion parameters has been released on Hugging Face, available at z-lab/Qwen3.6-27B-DFlash. This model likely implements speculative decoding or
A new Dflash drafting model based on Qwen 3.6 with 27 billion parameters has been released on Hugging Face, available at z-lab/Qwen3.6-27B-DFlash. This model likely implements speculative decoding or draft model techniques to improve inference efficiency. The announcement was made by Clem Delangue on X/Twitter.
Related
- Qwen3.6-27B-TQ3_4S is insanely good! https://huggingface.co/YTan2000/Qwen3.6-27B-TQ3_4S fit on my 16GB with 32k context Two prompts and I ge…
- Qwen3.6-35B-A3B is trending at #1 on Hugging Face! 🥇🤗 Thank you for making us the top trending model on @huggingface this week. Let's keep…
- DFlash for Kimi-K2.5 was pushed 3 hours ago! Acceptance length varies from 4.0 to 6.3 on datasets. This is only with SGLang btw. https://hug…
- The king reigns supreme This is likely going to be the best finetune of qwen 3.6 35b https://huggingface.co/DJLougen/Ornstein3.6-35B-A3B-GGU…
Source: Clem Delangue (X) | 2026-04-25