Model Releases
These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM t…
These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM tool 🚀 DeepSeek-V4 Preview is officially live & open-sourced!
These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM tool 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top closed-source models. 🔹 DeepSeek-V4-Flash: 284B total / 13B active params. Y…
Related
- 🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud.
- 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T t…
- [[ainews-deepseek-v4-pro-16t-a49b-and-flash-284b-a13b-base-and|[AINews] DeepSeek V4 Pro (1.6T-A49B) and Flash (284B-A13B), Base and Instruct — runnable on Huawei Ascend chips]]
- Built GPT-2, Llama 3, and DeepSeek from scratch in PyTorch - open source code + book [p]
Source: Simon Willison (X) | 2026-04-24