Model Card: Naive-N0.5-Flash
Naive-N0.5-Flash (HuggingFace: NaiveAI/Naive-N0.5-Flash): open-sourced 2026-09-27 under MIT, a 309B-total/15.5B-active MoE model built on Xiaomi's open MiMo-V2.5 base model with 3.25T further training tokens, replacing every full-attention layer with a hybrid of 39 Sliding-Window Attention layers and 9 DeepSeek Sparse Attention layers; its own NaiveRT inference runtime hits 50 tokens/s standard and up to 2,000 tokens/s in Ultrafast mode; NaiveAI's self-reported SWE-bench Pro score is 73.6 (behind Claude Opus 5.5's 89.9), and it trails same-tier open model DeepSeek V4.1 Flash on DeepSWE, Terminal-Bench, and ProgramBench; announced API pricing is $0.10 input / $0.40 output / $0.01 cache read per 1M tokens, but the API is not yet live; founder Dai Jifeng has an unresolved IP dispute with former employer MiroMind