Model Releases
SyzygyResearch/Mach-1-Additive-35B-GGUF · Hugging Face
They released both GGUFs & custom llama.cpp fork today. 35B MOE in 7GB size which's good for Mobile & Edge devices(Also low memory systems). Up to 120 t/s on Consumer Laptop. GGUFs: https://huggingfac
They released both GGUFs & custom llama.cpp fork today. 35B MOE in 7GB size which's good for Mobile & Edge devices(Also low memory systems). Up to 120 t/s on Consumer Laptop. GGUFs: https://huggingface.co/SyzygyResearch/Mach-1-Additive-35B-GGUF https://huggingface.co/SyzygyResearch/Mach-1-Additive-35B-Multimodal-GGUF Custom llama.cpp fork: https://github.com/SyzygyResearch/llama.cpp-mach1 Their 2 weeks old tweet below. Yes, Laguna S2.1 and Qwen 3.8 are on the way! BTW Track other similar models here : 1-bit / 2-bit / Ternary / Bitnet Models - Updates & Tracking submitted by /u/pmttyji [link] [comments]
Source: r/LocalLLaMA | 2026-08-20