Official Documentation Live Lineup:
Gemini 3Gemini 3.8 FlashGemini 3.8 LiveGemini 3.8 Live Extended ThinkingGemini 3.7 FlashGemini 3.6 FlashGemini 3.5 FlashGemini 3.5 Flash-Lite

Latest Releases & Architecture Updates

Ordered by release timestamp (driven by releases.json)
Live Multimodal Extended Reasoning 2026-09 Window: 2,000,000 Tokens

Gemini 3.8 Live Extended Thinking (Frontier Flagship)

Google's flagship Gemini 3.8 released September 2026, integrating native full-duplex live streaming with extended chain-of-thought reflection for complex real-time collaboration.

High-Throughput Scale Flagship 2026-09 Window: 1,000,000 Tokens

Gemini 3.8 Flash (Scale GA)

Google's high-throughput workhorse launched in early September 2026, cutting per-token latency and costs while preserving top reasoning for enterprise automation.

Ultra-Low Latency Efficient Tier 2026-08 Window: 500,000 Tokens

Gemini 3.5 Flash-Lite

Ultra-low-latency lightweight model engineered for sub-100ms simple chats and massive high-concurrency routing workflows.

Gemini Series Technical Specification Reference

Google AI Studio / Vertex AI standard configurations
Model Identifier Max Context Window Multimodal Capabilities Availability
Gemini 3.8 Live Extended Thinking 2M tokens Native millisecond full-duplex stream & chain-of-thought Latest Frontier
Gemini 3.8 Flash 1M tokens Enterprise high-throughput multimodal parsing General Availability
Gemini 3.5 Flash-Lite 500K tokens Ultra-low latency high-concurrency stream Production Lite