Model Releases
Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 suppo…
Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 support! @lightseekorg We’re proud to be the Day 0 open-source inf
Excited to see Qwen3.8 running at scale with TokenSpeed! 🚀 Light on latency, big on speed. Kudos to LightSeek for the fantastic Day-0 support! @lightseekorg We’re proud to be the Day 0 open-source inference engine partner for @Alibaba_Qwen 3.8. To serve this 2.4T-parameter model across multi-node @NVIDIAAI Blackwell inference, we optimized DP/EP scaling across nodes, delivering 30%+ faster performance than TP16, plus DSpark speculati…
Related
- From idea to implementation in one go.🏃♀️ Max-level intelligence, served fresh on Day 0. Qwen3.8-2.4T-A95B is live on SiliconFlow. Thanks!…
- Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️
- 📢Meet Qwen3.8-Max — our most capable model to date. Next week, the open weights of Qwen3.8-Max will be released, and Qwen3.8-27B is also go…
Source: Qwen (X) | 2026-08-14