GEMINI 3.5 FLASH
Launched at Google I/O 2026 (May 19). 76.2% Terminal-Bench, 83.6% MCP Atlas — ahead of Gemini 3.1 Pro at ~25% less cost. 1M context, vision, Google Search grounding. $1.50/$9 per million tokens.
Gemini 3.5 Flash launched at Google I/O 2026 on May 19 and broke the Flash line's usual positioning: instead of being the cheap sidekick to a Pro flagship, it beats Gemini 3.1 Pro on coding and agentic benchmarks — 76.2% on Terminal-Bench and 83.6% on MCP Atlas — while costing about 25% less ($1.50 per million input tokens and $9 per million output, vs the Pro's $2/$12). It keeps everything the Gemini line is known for: a 1M-token context window, native multimodal vision input, and googleSearch grounding for the freshest retrieval in the frontier set. Versus its Flash-line predecessor, Gemini 3 Flash ($0.50/$3), it costs 3× more but plays a different role — 3 Flash is a speed tier, 3.5 Flash is a frontier coding and agent model that happens to carry the Flash name. The case for it: agentic coding, tool-heavy workflows, and search-grounded work at a mid-tier price. The case against: tasks that lean on the Pro line's deeper reasoning, and bulk throughput where 3 Flash or 3.1 Flash-Lite remain far cheaper.
The Gemini lineup now brackets 3.5 Flash from both sides: Gemini 3.1 Pro ($2/$12) keeps the reasoning crown of the range, while Gemini 3 Flash ($0.50/$3) and 3.1 Flash-Lite ($0.25/$1.50) hold the throughput floor. 3.5 Flash occupies the newly interesting middle — Google's answer to 'frontier coding without flagship pricing'. If your workload is agentic or terminal-driven, start here rather than at Pro; if it's deep analysis or math, Pro still earns its premium; if it's high-volume and simple, drop down a tier.
Gemini 3.5 Flash is available on every paid Council AI plan (Plus, Pro, Ultra) and has become a default Google seat in cost-balanced councils: frontier-adjacent quality, search grounding, and a price that doesn't crowd out other voices in your monthly budget. Councils run it in parallel with models from up to 8 other labs, and the moderator scores agreement across the answers — its googleSearch grounding often makes it the fact-checking voice when other models disagree on current events.
$1.50 per million input tokens and $9.00 per million output tokens — about 25% less than Gemini 3.1 Pro ($2/$12). Council AI bundles it into monthly plan budgets.
1,000,000 tokens, in line with the rest of the modern Gemini range.
On coding and agentic benchmarks, yes — 76.2% Terminal-Bench and 83.6% MCP Atlas put it ahead of 3.1 Pro at lower cost. For deep reasoning and analysis workloads, 3.1 Pro remains the stronger pick.
May 19, 2026, launched at Google I/O 2026.
Yes — on all paid plans (Plus, Pro, Ultra). It runs in parallel councils with models from 8 other labs, with moderator synthesis and an agreement score on every run.