Model Card|Edge0-35B-A3B
Edge0-35B-A3B-preview (Edge0/Edge0-35B-A3B-preview): open-sourced by Edge0-AI on 2026-09-08. 4-bit quantized, 256 experts with 4 active per token (base model: Qwen3.5-MoE 35B-A3B). Three mechanisms — SSD expert offload, prerouter predictive routing, and Recover-LoRA distillation — let it run on a Mac mini M4 Pro (24GB) with just 2.9 GiB peak memory and 15 tok/s decode, averaging only a 3.9-point loss across 5 benchmarks versus the fp16 base. Fully open under Apache-2.0; also ships an 8B-A1B tier (built on Ling 3.0 Tiny). Agentic capability is currently weak — officially positioned as a preview release.