Model Releases
Help me complete my AI collection
I’m building the ultimate AI tool vault, but every great collection has a few missing pieces. Note: I will react to every comment AI's currently installed: Qwen3.5-0.8B-UD-Q4_K_XL.gguf(classification)
I’m building the ultimate AI tool vault, but every great collection has a few missing pieces. Note: I will react to every comment AI's currently installed: Qwen3.5-0.8B-UD-Q4_K_XL.gguf(classification) Qwen3.5-2B-UD-Q4_K_XL.gguf(Prompt enhancer, Routing, Approval ) Qwen3.5-4B-UD-Q4_K_XL.gguf(Instant) Qwen3.6-35B-A3B-UD-Q4_K_M.gguf(Quality) Qwen3.6-35B-A3B-Uncensored-Hauhau(test purposes) Qwen3-Coder-Next-UD-Q4_K_M.gguf(long horizon tasks) My Specs: GPU: RTX 5070 ti (16GB VRAM) RAM: Corsair vengeance 64GB 5200mt DDR5 CL40(dual-channel) CPU: intel i9 14900k SSD: Samsung s990 pro 2tb Backend: Llama.cpp server I tried GPT-OSS and was disappointed by tool usage. Gemma 4 was good and great tool usage but it was beaten by Qwen. Any recommendations?? Like something that you genuinely enjoyed or made you impressed. Feel free to share!! I will be reading every single comment. submitted by /u/Possible_Grocery8079 [link] [comments]
Related
- Extened garlic to run Qwen3.5 35B A3B float8 at 55 tok/s on RTX 5060 Ti
- Qwen 3.6 27B BF16 vs Q4_K_M vs Q8_0 GGUF evaluation
- OvisOCR2 (0.8B): first end-to-end model to top OmniDocBench - I threw 827 real scanned medical docs at it, here's everything I learned
- mudler/Qwen3.6-35B-A3B-Claude-4.7-Opus-Reasoning-Distilled-APEX-MTP-GGUF just released !
- Qwen3.6-35B-A3B tool calling benchmark: ByteShape vs. Unsloth GGUFs, KV cache quants & long context performance
Source: r/LocalLLaMA | 2026-07-25