NezhaGate
Chat

Gemini 3.8 Flash

gemini-3.8-flash
24H STATUS Success: 100.0%
-24hnow
Only successes and our own failures are counted: green is success in that hour, red is the failure share. Two things are excluded -- client-side errors (bad request, auth, insufficient balance and other 4xx) and requests still in flight. One bar per hour, taken from real gateway call logs.

Gemini 3.8 Flash is the latest Flash generation from Google (released September 2026), built for long-running coding, agent workflows and complex reasoning: 1M-token context, up to 64K output, image input. It uses a dynamic thinking budget -- quick answers for simple questions, more thinking for hard ones. Fully OpenAI-API compatible: set model to gemini-3.8-flash.

ChatLatest modelAdaptive thinking1M contextOpenAI compatible

Live Test · Playground

Try out Gemini 3.8 Flash right here (available after login)。

Input

Advanced
Web search
Memory · multi-turn

Conversation

Start a conversationType a message below to begin

About Gemini 3.8 Flash

Gemini 3.8 Flash is the latest Flash generation from Google (released September 2026), built for long-running coding, agent workflows and complex reasoning: 1M-token context, up to 64K output, image input, and a dynamic thinking budget that answers simple questions fast and thinks longer on hard ones. Fully OpenAI-API compatible: set model to gemini-3.8-flash. Priced below the Google list price on both input and output.

Use cases

Long-running coding and agent workflows

A 1M context holds a whole repository plus long conversation history, ideal for agents with many tool calls.

Batches of uneven difficulty

The dynamic thinking budget returns simple requests quickly and thinks more on hard ones, with no manual tiering.

Mixed image and text input

Image input is supported, so screenshot understanding and chart questions use the same endpoint.

How to choose

Pick 3.8 for the latest generation; stay on gemini-3.7-flash (same price) if your prompts are tuned for it; use the four gemini-3.6-flash tiers when you want manual control of the thinking budget.

FAQ

How does it differ from gemini-3.7-flash?
3.8 is the latest generation and 3.7 the previous one; same price, same dynamic thinking budget, with 3.8 aimed at longer-running coding and agent tasks.
Is it cheaper than the official price?
Yes. Both input and output are priced below the Google list price ($0.75 / $3.75 per 1M); the pricing page shows the comparison line by line.
Context and output limits?
1M-token context and up to 64K output; system prompts and other overhead count toward the window.
Am I charged for failed calls?
No. Only calls that return successfully are settled on actual usage; upstream errors and timeouts never touch your balance.

Related Models

Explore other models you can integrate.

View all →
GPT-5.6 Sol Chat

GPT-5.6 Sol

gpt-5.6-sol

Frontier flagship for agentic coding

$2.00/1M in · $12.00/1M out ↓75% View →
GPT-5.6 Terra Chat

GPT-5.6 Terra

gpt-5.6-terra

Balanced everyday agentic coding model

$1.20/1M in · $7.00/1M out ↓77% View →
GPT-5.6 Luna Chat

GPT-5.6 Luna

gpt-5.6-luna

Fast, economical agentic coding model

$0.80/1M in · $4.80/1M out ↓68% View →
GPT-5.5 Chat

GPT-5.5

gpt-5.5

Flagship chat and reasoning model

$0.70/1M in · $4.20/1M out ↓86% View →
GPT-6 Astra Chat

GPT-6 Astra

gpt-6-astra

Next-generation OpenAI flagship - 1.05M context

$2.80/1M in · $14.00/1M out ↓72% View →
GPT Image 2 Image

GPT Image 2

gpt-image-2

High-quality text-to-image / image-to-image model

$0.015/img and up View →
GPT Image 2.5 Flare Image

GPT Image 2.5 Flare

gpt-image-2.5-flare

Next-gen image model - clean and smooth

$0.015/img and up View →
GPT Image 2.5 Sunburst Image

GPT Image 2.5 Sunburst

gpt-image-2.5-sunburst

Next-gen image model - richer texture

$0.015/img and up View →
Nano Banana 2 Image

Nano Banana 2

nano-banana-2

High-quality text-to-image / image-to-image model

$0.025/img and up View →
Nano Banana Pro Image

Nano Banana Pro

nano-banana-pro

Flagship text-to-image / image-to-image model

$0.040/img and up View →
Claude Sonnet 4.6 Chat

Claude Sonnet 4.6

claude-sonnet-4-6

Balanced, efficient chat and coding model

$1.50/1M in · $7.50/1M out ↓50% View →
Claude Opus 5 Chat

Claude Opus 5

claude-opus-5

Next-generation flagship from Anthropic

$4.00/1M in · $20.00/1M out View →
Claude Fable 5 Chat

Claude Fable 5

claude-fable-5

Anthropic Fable series · narrative and long-form writing

$8.00/1M in · $40.00/1M out View →
Gemini 3.1 Pro Chat

Gemini 3.1 Pro

gemini-3.1-pro

Next-generation flagship reasoning model from Google

$0.50/1M in · $3.00/1M out ↓75% View →
Gemini 3.7 Flash Chat

Gemini 3.7 Flash

gemini-3.7-flash

Previous Flash · adaptive thinking

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.6 Flash Chat

Gemini 3.6 Flash

gemini-3.6-flash

Next-generation fast thinking model · four thinking budgets

$0.60/1M in · $3.60/1M out View →
Gemini 3.6 Flash High Chat

Gemini 3.6 Flash High

gemini-3.6-flash-high

Deep thinking tier

$0.60/1M in · $3.60/1M out View →
Gemini 3.6 Flash Low Chat

Gemini 3.6 Flash Low

gemini-3.6-flash-low

Fastest, lowest-cost tier

$0.60/1M in · $3.60/1M out View →
Gemini 3.6 Flash Tiered Chat

Gemini 3.6 Flash Tiered

gemini-3.6-flash-tiered

Auto-tiered thinking

$0.60/1M in · $3.60/1M out View →
Gemini 3 Flash Chat

Gemini 3 Flash

gemini-3-flash-preview

Fast, low-latency, cost-efficient chat model

$0.30/1M in · $1.20/1M out ↓57% View →
Gemini 3.5 Flash Chat

Gemini 3.5 Flash

gemini-3.5-flash

Next-generation fast thinking model

$0.45/1M in · $2.70/1M out ↓70% View →
Gemini 2.5 Flash Chat

Gemini 2.5 Flash

gemini-2.5-flash

Fast, low-latency, cost-efficient chat

$0.30/1M in · $1.20/1M out ↓46% View →
Veo 3.1 Video

Veo 3.1

veo-3.1

Text / image to video · multiple tiers · multiple resolutions

$0.075/clip+ View →
Seedance 2.5 Video

Seedance 2.5

seedance-2.5

Text / image to video · up to 30 seconds

$0.158/s View →
Seedance 2.0 Video

Seedance 2.0

seedance-2.0

Text / image to video · 5-15 seconds

$0.150/s View →
Seedance 2.0 Fast Video

Seedance 2.0 Fast

seedance-2.0-fast

Low latency · lowest cost

$0.100/s View →
Seedance 2.0 Mini Video

Seedance 2.0 Mini

seedance-2.0-mini

Entry tier · lowest cost

$0.066/s View →
Wan 3.0 Video

Wan 3.0

wan3.0-video

One endpoint, five modes · up to 30 seconds

$0.120/s View →
Wan 3.0 Prime Video

Wan 3.0 Prime

wan3.0-video-prime

Same capabilities · several times faster

$0.160/s View →
MiniMax H3 Video

MiniMax H3

minimax-h3

1080P · native audio

$0.036/s View →
Grok Imagine Video 1.5 Video

Grok Imagine Video 1.5

grok-imagine-video-1.5

Billed per clip · up to 15 seconds

$0.300/clip+ View →