NaiveAI released Naive-N0.5-Flash, a 309B MoE model optimized for coding and AI R&D, built using AI-centered development where models assist in their own design and training. The model features a 1M context window with hybrid attention architecture and achieves up to 2,000 tokens/s inference speed through the NaiveRT system. Weights and inference code are open-sourced under MIT license with API pricing available.