Model Releases
What's better than an open-weight multimodal model release? Well, the technical report. I just love how these labs like Qwen and DeepSeek co…
What's better than an open-weight multimodal model release? Well, the technical report. I just love how these labs like Qwen and DeepSeek continue to drop gem after gem. Qwen3.8-Flash is the latest in
What's better than an open-weight multimodal model release? Well, the technical report. I just love how these labs like Qwen and DeepSeek continue to drop gem after gem. Qwen3.8-Flash is the latest in efficient multimodal MoE models. Worth reading the report. ⚡Meet Qwen3.8-Flash, a multimodal MoE and an early preview of the Qwen4 architecture, now open-weight! The production version Qwen3.8-Flash will be available soon via QwenCloud API at just 0.16/1M input tokens and 0.47/1M output tokens. 125B parameters + 51B N-gram embeddings…
Related
- Brace yourselves. We just entered a new era of frontier multimodal models. DeepSeek-V4-Flash-Vision-Exp advances multimodal agent performanc…
- We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: …
- Gemma 4 Technical Report
Source: DAIR.AI (X) | 2026-08-26