Xiaomi released MiMo-V2.6 Pro and Flash under MIT license with its RL training stack open-sourced. Architecture, benchmarks, cost, and self-hosting notes.
Gemini 3.8 Flash launched Sept 2, 2026 with a 1M context window and introductory pricing that doubles on Jan 1, 2027. A cost-per-task breakdown for builders.
OpenAI shipped GPT-6 Sol and Luna on Sept 22, 2026 with a 1.05M context window and prices 50% below GPT-5.6. Here is what changed and how they compare.
Kimi K3 explained: Moonshot AI's 2.8T-parameter open-weight MoE - Kimi Delta Attention, 1M-token context, training, benchmarks, deployment cost and honest limits of the reasoning model in 2026.
Physical Intelligence pi0.5 explained: the ~3.3B vision-language-action robot foundation model - architecture, FAST action tokenization, training on 10+ embodiments, benchmarks, deployment and honest limits in 2026.
Mistral Large 3 explained: the 675B/41B Apache-2.0 sparse MoE with a 262K context window - architecture, training, benchmarks, licensing, self-hosting VRAM, pricing and honest limits in 2026.