Skip to main content
The Human Bit

Google · Gemini

Gemini 3.7 Flash

Google's prior fastest Flash model for complex agentic tasks at scale, now sharing the Flash tier with Gemini 3.8 Flash, which Google names as the latest and most capable Flash model.

The Human Bit position

Treat this as a migration entry. If you are choosing a Gemini Flash model today the comparison is 3.8 Flash against 3.7 Flash on your own task, since both currently carry the same published rate.

Use it for

  • Existing integrations not yet migrated to 3.8 Flash
  • Complex agentic tasks and multi-step coding loops
  • Fast multimodal document and code work
Leave it alone when

Do not treat the current rate as permanent: Google's own pricing page schedules a step-up on 1 January 2027, so a workflow costed today should be re-checked before that date. A new build should also compare it against 3.8 Flash rather than start here.

Effort without the jargon

Default

You need a broadly capable, fast Gemini model and will validate the output in its final workflow.

Advanced recordCost, limits, privacy and operations

This is a fast, current-generation Gemini for coding and agentic work, and its published rate today matches 3.6 Flash exactly; both are scheduled to roughly double on 1 January 2027, so do not plan a durable budget off today's number alone.

API cost
Officially disclosedPaid standard API rate: US$0.75 per million input tokens and US$3.75 per million output tokens including thinking, through 31 December 2026, rising to US$1.50 and US$7.50 from 1 January 2027, identical to 3.6 Flash's published rate. Context caching is US$0.075 per million tokens (US$0.15 from January 2027) plus US$0.50 per million tokens per hour of storage (US$1.00 from January 2027).
Context window
Officially disclosed1,048,576 input tokens
Maximum output
Officially disclosed65,536 output tokens
Knowledge cutoff
Not disclosedNot stated on the model card
Recorded inputs
Text · Images · Video · Audio · PDF
Recorded output
Text

Data handling is a product decision

Depends on surfaceData-use conditions differ by free and paid Gemini API tiers and by Google product surface; verify the live pricing and service terms for the route you use.

Operational constraints

  • Standard-tier input and output pricing is scheduled to increase on 1 January 2027; re-check the live rate before treating today's price as durable.
  • The published rate is currently identical to 3.6 Flash's, so price alone is not a reason to migrate.
  • The Gemini app, AI Mode in Search, AI Studio, the Gemini API and Gemini Enterprise are different operational surfaces.
  • Grounding, media inputs and cache storage can add charges beyond input and output generation.

Unknown means the cited official record does not disclose a safe value. It is not an estimate. Prices are provider list prices in USD where stated and can change before this record's review date.

Keep the evidence separate

Canonical provider claim

Google's Flash overview page now leads with Gemini 3.8 Flash as its most intelligent workhorse model for coding and agents; 3.7 Flash remains available with the pricing, limits and January 2027 rate change originally documented for it.

Independent observation

No sufficiently useful independent note has been added yet. That absence is not filled with our guess.

Readiness · informational

No Human Bit result is claimed; use this as evaluation guidance only.

Human Bit test

Not scheduled

No Human Bit result is claimed for this model yet.

Sources