Syzygy Releases Mach-1 35B
·2026.08.20 16:52
Key point
SyzygyResearch has released the 7GB Mach-1-Additive-35B model and a dedicated llama.cpp fork.
Details
SyzygyResearch has released the GGUF version of the Mach-1-Additive-35B model and a custom llama.cpp fork on Hugging Face and GitHub.
This model compresses a 35B MoE architecture into 7GB, optimized for execution on mobile devices, edge devices, and low-memory systems. It has been confirmed to achieve inference speeds of up to 120 t/s in consumer laptop environments.
- Model: Mach-1-Additive-35B (GGUF)
- Size: 7GB
- Inference Speed: Up to 120 t/s (on consumer laptops)
- Distribution: Hugging Face and GitHub (SyzygyResearch/llama.cpp-mach1)
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.