Google · Gemini
Gemini 3.5 Flash
Google’s earlier fast Flash model, still stable and deployed broadly across consumer, developer and enterprise products, but no longer the current Flash: Google now points new work at 3.6, 3.7 or 3.8 Flash.
A serious candidate for Google-heavy work. Its first comparison should include speed, total cost and the quality of Workspace handoffs.
Use it for
- Fast multimodal document work
- Agentic and coding workflows
- Google-connected products and services
Do not read the Flash name as automatically cheap; measure token volume and total task cost against the prior workflow.
Effort without the jargon
You need a broadly capable, fast Gemini model and will validate the output in its final workflow.
Advanced recordCost, limits, privacy and operations
Flash signals speed, not a guaranteed cheap task; check the current API rate card and measure thinking, media, grounding and caching charges together.
- API cost
- Officially disclosedPaid standard API rate: US$1.50 per million input tokens and US$9.00 per million output tokens including thinking. That is six times the Flash-Lite output rate, so the Flash name is not a price signal; caching, grounding and media modes are charged separately.
- Context window
- Officially disclosed1,048,576 input tokens
- Maximum output
- Officially disclosed65,536 output tokens
- Knowledge cutoff
- Not disclosedNot stated on the model card
- Recorded inputs
- Text · Images · Video · Audio · PDF
- Recorded output
- Text
Data handling is a product decision
Depends on surfaceData-use conditions differ by free and paid Gemini API tiers and by Google product surface; verify the live pricing and service terms for the route you use.
Operational constraints
- The Gemini app, AI Mode in Search, AI Studio, the Gemini API and Gemini Enterprise are different operational surfaces.
- Grounding, media inputs and cache storage can add charges beyond input and output generation.
- Audio generation, image generation and the Live API are not supported; computer use is preview only.
Unknown means the cited official record does not disclose a safe value. It is not an estimate. Prices are provider list prices in USD where stated and can change before this record's review date.
Keep the evidence separate
Google’s launch material positioned 3.5 Flash as frontier-level intelligence at Flash speed. Its current model list now calls 3.5 Flash the legacy Flash model for routine, high-throughput workloads and names Gemini 3.8 Flash as the latest and most capable Flash model.
He found the launch technically broad but highlighted a substantial effective price increase and the need to measure total benchmark or task cost, not the Flash label.
Read the independent source ↗No Human Bit result is claimed; use this as evaluation guidance only.
Not scheduled
No Human Bit result is claimed for this model yet.