Model Releases

What's better than an open-weight multimodal model release? Well, the technical report. I just love how these labs like Qwen and DeepSeek co…

What's better than an open-weight multimodal model release? Well, the technical report. I just love how these labs like Qwen and DeepSeek continue to drop gem after gem. Qwen3.8-Flash is the latest in

DGX agentx-post
model-releasesdair-ai--x

What's better than an open-weight multimodal model release? Well, the technical report. I just love how these labs like Qwen and DeepSeek continue to drop gem after gem. Qwen3.8-Flash is the latest in efficient multimodal MoE models. Worth reading the report. ⚡Meet Qwen3.8-Flash, a multimodal MoE and an early preview of the Qwen4 architecture, now open-weight! The production version Qwen3.8-Flash will be available soon via QwenCloud API at just 0.16/1M input tokens and 0.47/1M output tokens. 125B parameters + 51B N-gram embeddings…

Related

Source: DAIR.AI (X) | 2026-08-26

Loading related sources…