Gemini 3.8 Flash
⧉Gemini 3.8 Flash is the latest Flash generation from Google (released September 2026), built for long-running coding, agent workflows and complex reasoning: 1M-token context, up to 64K output, image input. It uses a dynamic thinking budget -- quick answers for simple questions, more thinking for hard ones. Fully OpenAI-API compatible: set model to gemini-3.8-flash.
Live Test · Playground
Try out Gemini 3.8 Flash right here (available after login)。
Input
Advanced
Conversation
About Gemini 3.8 Flash
Gemini 3.8 Flash is the latest Flash generation from Google (released September 2026), built for long-running coding, agent workflows and complex reasoning: 1M-token context, up to 64K output, image input, and a dynamic thinking budget that answers simple questions fast and thinks longer on hard ones. Fully OpenAI-API compatible: set model to gemini-3.8-flash. Priced below the Google list price on both input and output.
Use cases
A 1M context holds a whole repository plus long conversation history, ideal for agents with many tool calls.
The dynamic thinking budget returns simple requests quickly and thinks more on hard ones, with no manual tiering.
Image input is supported, so screenshot understanding and chart questions use the same endpoint.
How to choose
Pick 3.8 for the latest generation; stay on gemini-3.7-flash (same price) if your prompts are tuned for it; use the four gemini-3.6-flash tiers when you want manual control of the thinking budget.
FAQ
How does it differ from gemini-3.7-flash?
Is it cheaper than the official price?
Context and output limits?
Am I charged for failed calls?
Related Models
Explore other models you can integrate.
GPT-5.6 Sol
Frontier flagship for agentic coding
GPT-5.6 Terra
Balanced everyday agentic coding model
GPT-5.6 Luna
Fast, economical agentic coding model
GPT-5.5
Flagship chat and reasoning model
GPT-6 Astra
Next-generation OpenAI flagship - 1.05M context
GPT Image 2
High-quality text-to-image / image-to-image model
GPT Image 2.5 Flare
Next-gen image model - clean and smooth
GPT Image 2.5 Sunburst
Next-gen image model - richer texture
Nano Banana 2
High-quality text-to-image / image-to-image model
Nano Banana Pro
Flagship text-to-image / image-to-image model
Claude Sonnet 4.6
Balanced, efficient chat and coding model
Claude Opus 5
Next-generation flagship from Anthropic
Claude Fable 5
Anthropic Fable series · narrative and long-form writing
Gemini 3.1 Pro
Next-generation flagship reasoning model from Google
Gemini 3.7 Flash
Previous Flash · adaptive thinking
Gemini 3.6 Flash
Next-generation fast thinking model · four thinking budgets
Gemini 3.6 Flash High
Deep thinking tier
Gemini 3.6 Flash Low
Fastest, lowest-cost tier
Gemini 3.6 Flash Tiered
Auto-tiered thinking
Gemini 3 Flash
Fast, low-latency, cost-efficient chat model
Gemini 3.5 Flash
Next-generation fast thinking model
Gemini 2.5 Flash
Fast, low-latency, cost-efficient chat
Veo 3.1
Text / image to video · multiple tiers · multiple resolutions
Seedance 2.5
Text / image to video · up to 30 seconds
Seedance 2.0
Text / image to video · 5-15 seconds
Seedance 2.0 Fast
Low latency · lowest cost
Seedance 2.0 Mini
Entry tier · lowest cost
Wan 3.0
One endpoint, five modes · up to 30 seconds
Wan 3.0 Prime
Same capabilities · several times faster
MiniMax H3
1080P · native audio
Grok Imagine Video 1.5
Billed per clip · up to 15 seconds