Google · Gemini 3.5 Flash · May 2026

Gemini 3.5 Flash — near-Pro ability at a faster pace

  • Million-token context
  • Multimodal input
  • Adjustable thinking depth
  • Parallel agent loops
Gemini 3.5 Flash

What is Gemini 3.5 Flash?

Gemini 3.5 Flash is Google’s May 2026 release aimed at near-Pro coding and reasoning while keeping faster replies. Versus earlier Gemini 3 Flash Preview, agent loops and thinking controls are more complete. For new similar work, evaluate Gemini 3.6 Flash first; for a higher ceiling, look at Gemini 3.1 Pro. On iMini Agent, pick Gemini 3.5 Flash to use it.

Vendor
Google
Released
May 2026
Context
1M tokens
Max output
~64K–65K
Deep thinking
Usually medium · adjustable
Best as
Proven workflows

What’s new versus Gemini 3 Flash Preview

Near-Pro coding, thinking levels, and multimodal agents—the capabilities to check before you stay on 3.5.

Coding and reasoning closer to Pro

Google positions 3.5 Flash as near-Pro coding and reasoning while keeping faster replies and a leaner cost posture.

More complete thinking depth

Tune from lighter to deeper thinking by task; medium depth usually works for both everyday and slightly harder jobs on one model.

Parallel agent loops

Better suited to multi-tool, multi-step workflows that run in parallel—and can keep history in long context.

Full multimodal input

Text, images, video, audio, and PDFs can enter the same task—suited to mixed UI, document, and AV materials.

Official evaluations

Artificial Analysis results cited in Google’s Gemini 3.5 launch materials: intelligence index and output speed on one scatter plot—showing where Flash sits between ability and throughput.

Artificial Analysis: intelligence index vs output speed, highlighting Gemini 3.5 Flash
Artificial Analysis intelligence index vs output speed (data as of 2026-05-13). Gemini 3.5 Flash scores ~57 on the index at ~285 tokens/s—same intelligence band as Claude Opus 4.7 and Gemini 3.1 Pro, with clearly higher throughput. That is the core evidence for the Flash positioning.

Three common workflows

Gemini 3.5 Flash still fits proven workflows; for new picks, compare 3.6.

Near-Pro everyday coding

For feature work and medium-difficulty fixes. Stay on 3.5 if the workflow is locked; try 3.6 for new work.

Long-context agents

For automation that carries long specs and tool traces. Agree on completion criteria up front.

Multimodal material organization

For understanding and summarizing screenshots, PDFs, and AV materials.

How to choose vs Gemini 3.6 Flash and Gemini 3.1 Pro

All three are Gemini 3. The gap is mainly generation efficiency and the Pro ceiling.

CriterionGemini 3.5 FlashGemini 3.6 FlashGemini 3.1 Pro
Context1M tokens1M tokens1M tokens
ReasoningUsually medium · adjustableAdjustable thinking, leaner outputPro-tier reasoning ceiling
Coding & agentsNear-Pro fast codingFaster, leaner codingHeavier agents & complex workflows
Long-horizon toolsParallel agent loopsHigh-turnaround, more restrainedFavors reliability
Speed postureFaster repliesFaster and leanerFavors precision
Prefer whenWorkflows still locked on 3.5New fast coding & agent workYou need higher precision

Using Gemini 3.5 Flash on iMini

Free quota, model naming, key specs, and how it differs from 3.6 and Pro.

Use Gemini 3.5 Flash free on iMini Agent

Million-token context, near-Pro ability, faster replies.

Open Gemini 3.5 Flash

No Google API key required.