NaiveAI open-sources Naive-N0.5-Flash, a 309B MoE coding model with native 1M context under MIT
Built for coding and AI R&D on Xiaomi's MiMo-V2.5-Base, the model activates 15.5B parameters per token and reaches 1M tokens of context with sliding-window plus DeepSeek Sparse Attention and no full-attention layers; weights and inference code are MIT. NaiveAI also plans an API at $0.10 in / $0.40 out per million tokens, but self-hosting the full 49-shard checkpoint takes roughly 690 GB of memory.