Skip to main content
The Human Bit

Google · Gemini

Gemini 3.8 Flash

Google's current fastest Flash model for complex agentic tasks at scale, positioned as the successor to 3.7 Flash with stronger long-horizon software engineering and multi-step reasoning, while currently sharing the same published rate.

The Human Bit position

A serious default for coding-heavy Gemini work today. Because its published rate currently matches 3.7 Flash, the first real test should be whether the long-horizon coding and reasoning gains show up on your own task, not the price.

Use it for

  • Long-horizon software engineering and multi-step coding loops
  • Multi-step reasoning in specialised domains such as finance and legal work
  • Fast multimodal document and code work
Leave it alone when

Do not treat the current rate as permanent: Google's own pricing page schedules a step-up on 1 January 2027, the same date already set for 3.7 Flash, so a workflow costed today should be re-checked before that date.

Effort without the jargon

Default

You need a broadly capable, fast Gemini model and will validate the output in its final workflow.

Advanced recordCost, limits, privacy and operations

This is a fast, current-generation Gemini for coding and agentic work, and its published rate today matches 3.7 Flash exactly; both are scheduled to roughly double on 1 January 2027, so do not plan a durable budget off today's number alone.

API cost
Officially disclosedPaid standard API rate: US$0.75 per million input tokens and US$3.75 per million output tokens including thinking, through 31 December 2026, rising to US$1.50 and US$7.50 from 1 January 2027, identical to 3.7 Flash's published rate.
Context window
Officially disclosed1,048,576 input tokens
Maximum output
Officially disclosed65,536 output tokens
Knowledge cutoff
Not disclosedNot stated on the model card
Recorded inputs
Text · Images · Video · Audio · PDF
Recorded output
Text

Data handling is a product decision

Depends on surfaceData-use conditions differ by free and paid Gemini API tiers and by Google product surface; verify the live pricing and service terms for the route you use.

Operational constraints

  • Standard-tier input and output pricing is scheduled to increase on 1 January 2027; re-check the live rate before treating today's price as durable.
  • The published rate is currently identical to 3.7 Flash's, so price alone is not a reason to migrate.
  • Gemini 3.8 Flash Cyber is a separate, access-restricted variant through the Fairwind Program, not something available on the standard API.
  • The Gemini app, AI Mode in Search, AI Studio, the Gemini API and Gemini Enterprise are different operational surfaces.

Unknown means the cited official record does not disclose a safe value. It is not an estimate. Prices are provider list prices in USD where stated and can change before this record's review date.

Keep the evidence separate

Canonical provider claim

Google calls Gemini 3.8 Flash its most intelligent workhorse model yet for coding and agents, reporting significant gains over 3.7 Flash on long-horizon software engineering benchmarks and on multi-step reasoning in specialised domains. It also introduced Gemini 3.8 Flash Cyber, a cybersecurity-specialised variant restricted to its Fairwind Program for vetted government, critical-infrastructure and software-maintainer defenders.

Independent observation

No sufficiently useful independent note has been added yet. That absence is not filled with our guess.

Readiness · informational

No Human Bit result is claimed; use this as evaluation guidance only.

Human Bit test

Not scheduled

No Human Bit result is claimed for this model yet.

Sources