Google has quietly made two new Gemini models—Gemini 3.6 Flash and Gemini 3.5 Flash Lite—available through Google AI Studio and the Vertex AI API, expanding its speed-focused Flash lineup without a formal launch event.
July 21, 2026 · Google Gemini
Google quietly ships two new Flash models — no launch event
Gemini 3.6 Flash and Gemini 3.5 Flash Lite appeared in AI Studio and the Vertex AI API, selectable by developers with no headline announcement — expanding Google's speed-focused Flash tier as its next-gen Pro model stays delayed.
Gemini 3.6 Flash
Capability tier. "Most intelligent model yet for sustained frontier performance in agentic and coding tasks."
Gemini 3.5 Flash Lite
Scale tier. "High-throughput, low-latency execution for scaling high-volume agentic tasks and subagent workflows."
Reference pricing — Gemini 3.5 Flash (per million tokens, AI Studio)
Output runs 6× the cost of input — the two new models' own pricing is not yet published.
2
new Flash models, one two-way split
~45%
speed gain cited on a prior Lite release
0
official benchmarks or context specs released
Developers welcome it
Immediate availability drew appreciation; screenshots of the models running in AI Studio circulated fast among builders eager to test.
Skeptics push back
Some question the naming and cadence, reading the heavy focus on Flash as a sign the higher-end Pro release is behind schedule.
The takeaway for builders
A two-model split aimed squarely at agent apps — a capable model for coding and sustained reasoning, and a lightweight one tuned for high-volume execution and subagent orchestration. Real gains hinge on pricing, context limits and benchmarks Google has yet to disclose.
Continue reading The rest of this article is for AI News Blitz readers. Choose an option below to keep reading.
Already purchased? Sign in ✓ Signed in — this article isn’t included in your current plan.Unlocking the full article…