Frontier lab releases, open-source checkpoints, multimodal systems, inference stacks, and model capability shifts.
Google ships Gemini 3.6 Flash as latest frontier-tier fast model
OpenGoogle’s most recent tracked release is **Gemini 3.6 Flash**, a frontier-capable but latency-optimized model that extends the Gemini 3.x fast tier.[1][3] Gemini Flash is positioned as a high‑throughput multimodal model, maintaining strong agentic and coding performance at lower cost per token.[1][3]
OpenAI GPT‑5.6 family (Sol/Terra/Luna) becomes new flagship frontier stack
OpenOpenAI’s **GPT‑5.6** family (Sol, Terra, Luna) is now generally available, with Sol as the frontier flagship, Terra delivering ~GPT‑5.5 quality at half the cost, and Luna as the small, fast tier.[2][5][14] All three tiers share roughly 1M‑token context, updated coding and computer‑use capabilities, and a new Ultra sub‑agent mode with Max reasoning‑effort for the highest‑tier Sol configuration.[2][5][14]
Alibaba Qwen 3.8‑Max‑Preview and Moonshot Kimi K3 push open‑weight frontier models
OpenDemandSphere’s frontier tracker added **Qwen 3.8‑Max‑Preview** (2.4T parameters, Qwen’s first multimodal model above 1T) and **Kimi K3** (2.8T‑parameter open‑weight multimodal MoE with 1M context) to its July changelog.[8] These models extend the rapidly growing set of large open‑weight systems that approach closed‑lab frontier capabilities while remaining callable under commercial‑friendly licenses.[8][10]
