GEMINI 3.5 FLASH

Gemini 3.5 Flash: the Flash that beats the Pro

Launched at Google I/O 2026 (May 19). 76.2% Terminal-Bench, 83.6% MCP Atlas — ahead of Gemini 3.1 Pro at ~25% less cost. 1M context, vision, Google Search grounding. $1.50/$9 per million tokens.

Try Gemini 3.5 Flash on Council See it in a council

Gemini 3.5 Flash launched at Google I/O 2026 on May 19 and broke the Flash line's usual positioning: instead of being the cheap sidekick to a Pro flagship, it beats Gemini 3.1 Pro on coding and agentic benchmarks — 76.2% on Terminal-Bench and 83.6% on MCP Atlas — while costing about 25% less ($1.50 per million input tokens and $9 per million output, vs the Pro's $2/$12). It keeps everything the Gemini line is known for: a 1M-token context window, native multimodal vision input, and googleSearch grounding for the freshest retrieval in the frontier set. Versus its Flash-line predecessor, Gemini 3 Flash ($0.50/$3), it costs 3× more but plays a different role — 3 Flash is a speed tier, 3.5 Flash is a frontier coding and agent model that happens to carry the Flash name. The case for it: agentic coding, tool-heavy workflows, and search-grounded work at a mid-tier price. The case against: tasks that lean on the Pro line's deeper reasoning, and bulk throughput where 3 Flash or 3.1 Flash-Lite remain far cheaper.

Specs

ProviderGoogle DeepMind
ReleasedMay 19, 2026 (Google I/O 2026)
Context window1,000,000 tokens
Terminal-Bench76.2% — ahead of Gemini 3.1 Pro
MCP Atlas83.6%
VisionText + image input
Web searchNative googleSearch grounding
Price (in / out per 1M)$1.50 / $9.00

Best at

Where it loses

Where it sits in the Gemini range

The Gemini lineup now brackets 3.5 Flash from both sides: Gemini 3.1 Pro ($2/$12) keeps the reasoning crown of the range, while Gemini 3 Flash ($0.50/$3) and 3.1 Flash-Lite ($0.25/$1.50) hold the throughput floor. 3.5 Flash occupies the newly interesting middle — Google's answer to 'frontier coding without flagship pricing'. If your workload is agentic or terminal-driven, start here rather than at Pro; if it's deep analysis or math, Pro still earns its premium; if it's high-volume and simple, drop down a tier.

How Council AI runs it

Gemini 3.5 Flash is available on every paid Council AI plan (Plus, Pro, Ultra) and has become a default Google seat in cost-balanced councils: frontier-adjacent quality, search grounding, and a price that doesn't crowd out other voices in your monthly budget. Councils run it in parallel with models from up to 8 other labs, and the moderator scores agreement across the answers — its googleSearch grounding often makes it the fact-checking voice when other models disagree on current events.

Frequently asked questions

How much does Gemini 3.5 Flash cost via API?

$1.50 per million input tokens and $9.00 per million output tokens — about 25% less than Gemini 3.1 Pro ($2/$12). Council AI bundles it into monthly plan budgets.

What is Gemini 3.5 Flash's context window?

1,000,000 tokens, in line with the rest of the modern Gemini range.

Is Gemini 3.5 Flash really better than Gemini 3.1 Pro?

On coding and agentic benchmarks, yes — 76.2% Terminal-Bench and 83.6% MCP Atlas put it ahead of 3.1 Pro at lower cost. For deep reasoning and analysis workloads, 3.1 Pro remains the stronger pick.

When was Gemini 3.5 Flash released?

May 19, 2026, launched at Google I/O 2026.

Is Gemini 3.5 Flash available in Council AI?

Yes — on all paid plans (Plus, Pro, Ultra). It runs in parallel councils with models from 8 other labs, with moderator synthesis and an agreement score on every run.