All models

Google

Gemini 3.8 Flash

Google's stable, high-efficiency multimodal model for agents, software work, and large mixed-media inputs at an introductory Flash-tier price.

Plain-English overview

What Gemini 3.8 Flash actually is

Gemini 3.8 Flash combines a million-token input window with text, image, video, audio, and PDF understanding. Its native API features include function calling, code execution, search grounding, file search, and structured output.

The current price is promotional. Teams comparing annual costs should model both the introductory rate and the published higher rate that begins in 2027, plus any search or grounding calls.

Good fit for

  • High-volume multimodal agents
  • Long video, audio, PDF, and code inputs
  • Teams optimizing capability per token cost

Category comparison

The facts that matter for text models

These are provider-published specifications, not Cody benchmark scores. Follow the linked sources for current limits and endpoint-specific exceptions.

Context windowMaximum combined prompt and working context documented by the provider.
1,048,576 tokens
Maximum outputProvider-published response limit, where available.
65,536 tokens
Input priceCurrent standard list price per million input tokens unless noted.
$0.75 / 1M promo
Output priceCurrent standard list price per million output tokens unless noted.
$3.75 / 1M promo
InputsMedia types accepted by the listed model endpoint.
Text, image, video, audio, PDF
Tools & agentsSelected native tools and agent-building capabilities, not an exhaustive list.
Functions, search, code execution, file search

API and provider access

Where to get Gemini 3.8 Flash

Availability

Regions and access stage

Available through the Gemini API and Google AI Studio in supported regions, with Vertex and gateway routes also available.

Use Google's live available-regions list and verify feature restrictions for the selected route.

Check live availability

Data and training

The route matters.

Google says paid Gemini API prompts and responses are not used to improve its products. Unpaid-service content generally may be used, with different treatment for EEA, UK, and Swiss users; check the current terms for your account and region.

This is a concise reading of the cited provider material, not legal advice. A third-party gateway can have different storage, routing, training, and residency terms from the model maker's direct API.

Read the provider policy

Frequently asked questions

Gemini 3.8 Flash FAQ

What is Gemini 3.8 Flash?

Google's stable, high-efficiency multimodal model for agents, software work, and large mixed-media inputs at an introductory Flash-tier price. Gemini 3.8 Flash combines a million-token input window with text, image, video, audio, and PDF understanding. Its native API features include function calling, code execution, search grounding, file search, and structured output.

When was Gemini 3.8 Flash released?

Gemini 3.8 Flash was released on September 2, 2026 according to the cited provider materials.

Where can I access Gemini 3.8 Flash?

Available through the Gemini API and Google AI Studio in supported regions, with Vertex and gateway routes also available. The access routes listed in this guide are Google, OpenRouter.

How much does Gemini 3.8 Flash cost?

$0.75 input · $3.75 output / 1M tokens. Introductory paid-tier price through December 31, 2026; Google lists $1.50 input and $7.50 output beginning January 1, 2027.

Where is Gemini 3.8 Flash available?

Available through the Gemini API and Google AI Studio in supported regions, with Vertex and gateway routes also available. Use Google's live available-regions list and verify feature restrictions for the selected route.

Is my Gemini 3.8 Flash API data used for training?

Google says paid Gemini API prompts and responses are not used to improve its products. Unpaid-service content generally may be used, with different treatment for EEA, UK, and Swiss users; check the current terms for your account and region. The policy belongs to the provider route and account terms, so verify it again before production use.