Model Releases
We could really use Qwen3.8 in 27B, 35B, 122B and 397B sizes
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even dream of running the recent 1.5-2T+ beast
Instead of 2T+ models, continuing to release highly capable small to medium size LLMs would really help to keep this community vibrant. Hardly anyone can even dream of running the recent 1.5-2T+ beasts, while the range from the title could run comfortably (especially with CPU expert offloading) across a wide range of systems we have today. The trend towards Chinese labs trying to match the Mythos class frontier with trillion parameter open weights models is not helping the local model community to innovate. It just gives big corporates who can actually run these a cheaper alternative to the commercial frontier. submitted by /u/Responsible_Fig_1271 [link] [comments]
Related
- If anyone is running qwen 9b or 27b or 35b and getting wrong facts while web search, follow this.
- 90 agentic bakeoff runs: ThinkingCap vs Fable Fusion vs stock Qwen3.6-27B
- mudler/Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-Distilled-APEX-MTP-GGUF just released !
- Will small model intelligence be limited by parameter count?
- BeeLlama.cpp: advanced DFlash & TurboQuant with support of reasoning and vision. Qwen 3.6 27B Q5 with 200k context on 3090, 2-3x faster than baseline (peak 135 tps!)
Source: r/LocalLLaMA | 2026-07-27