Google (Gemini) · Gemini 3.8 Flash
gemini-3.8-flashAvailable from the provider.
- Released
- Deprecatednot announced
- Retirednot announced
Summary
Gemini 3.8 Flash is the newest model in Google's Flash line, at $0.75 per million input tokens and $3.75 per million output tokens. It takes text, image, video, audio and PDF input within a 1,048,576-token context window and returns up to 65,536 output tokens.
Google positions Gemini 3.8 Flash for long-horizon software engineering, autonomous agents and complex enterprise workflows at Flash-level speed and cost. It is the most recent Flash release, listed with a September 2026 update, and it carries the same $0.75 / $3.75 price as Gemini 3.7 Flash and Gemini 3.6 Flash.
It supports structured outputs, context caching, function calling, code execution, search grounding, URL context, file search, the Batch API and thinking at low, medium or high. Computer use is available in preview. There is no Live API support and no image or audio generation; the output is text.
No deprecation has been announced.
Tags
Editorial tags reflect our assessment; documented tags are backed by the sources below.
Tier
Mid-tierAccess
Hosted APISpecifications
- Input price
- $0.75 / M tok
- Output price
- $3.75 / M tok
- Price class
- Standard
- Context window
- 1M tokens
- Max output
- 65,536 tokens
- Released
- 2 September 2026
- Input modalities
- text, image
- Output modalities
- text
- Knowledge cutoff
- unknown
- Tool use
- Yes
- Structured output
- Yes
- License
- unknown
- Maker
- Google (Gemini)
Use it for
- Run agentic coding and long-running tool loops where a Flash-class price makes many calls affordable.
- Process mixed input in one request, combining text with images, video, audio or PDFs.
- Ground answers in live search results or URL content without a separate retrieval stack.
- Handle inputs up to roughly 1M tokens without chunking.
Avoid it when
- You need more than 65,536 output tokens in a single response; that limit is half of what the Claude and GPT flagships allow.
- You need a live, low-latency voice or streaming session; the Live API is not supported on this model.
- You need image or audio generation, which this model does not do.
- Cost per call is the binding constraint; Gemini 3.5 Flash-Lite costs $0.3 / $2.5 and Gemini 2.5 Flash-Lite $0.1 / $0.4.
Alternatives
What to use instead. Curated suggestions are written by our editors for this model; derived ones are ranked automatically from tags, tier and price class.
- Gemini 3.7 Flashcurated
Google (Gemini)
gemini-3.7-flashThe previous Flash release at the same $0.75 / $3.75 and the same limits, if you want to change one generation at a time.
In $0.75 / M tokOut $3.75 / M tok
ActiveTextVisionChat+6 - Gemini 3.5 Flash-Litecurated
Google (Gemini)
gemini-3.5-flash-liteLess than half the input price at $0.3 / $2.5 with the same modalities, for high-throughput or sub-agent work.
In $0.30 / M tokOut $2.50 / M tok
ActiveTextVisionChat
Change history
Nothing has changed since this model was first recorded — no status change and no price change.
Price history
No price changes recorded yet.
Sources & verification
Facts
Status, dates, pricing and specifications were verified against these pages.
Last verified: (2026-09-10 12:00) · ✓ verified
Editorial profile
The summary, the advice and the curated alternatives are our assessment, not a provider claim.
Editorial profile reviewed on by llm-researcher.