🚀 New models: GPT-6 Astra (1.05M-token context) and GPT Image 2.5 Flare / Sunburst are live →
NezhaGate
OpenAI-Compatible AI Gateway

Chat, images and video,through one API

NezhaGate is an OpenAI-compatible AI model gateway that helps developers connect to GPT, Claude, Gemini and image-generation models through one API. Built-in API key management, balance and usage, a model market and developer docs — billed by usage, with no charge on failed requests.

OpenAI SDK compatibleStreaming outputPay-as-you-go billingNo charge on failure
$0No charge for failed tasks
OpenAIAPI format compatible
Meteredbilling · no hidden fees
21open models · always growing

Model market

Top models hand-picked across every category — start integrating in minutes.

View all →
GPT-5.6 Terra Chat

GPT-5.6 Terra

gpt-5.6-terra

Balanced everyday agentic coding model

$1.20/1M in · $7.00/1M out ↓77% View →
GPT-5.5 Chat

GPT-5.5

gpt-5.5

Flagship chat and reasoning model

$0.70/1M in · $4.20/1M out ↓86% View →
GPT-6 Astra Chat

GPT-6 Astra

gpt-6-astra

Next-generation OpenAI flagship - 1.05M context

$2.80/1M in · $14.00/1M out ↓72% View →
GPT Image 2 Image

GPT Image 2

gpt-image-2

High-quality text-to-image / image-to-image model

$0.015/img and up View →
GPT Image 2.5 Flare Image

GPT Image 2.5 Flare

gpt-image-2.5-flare

Next-gen image model - clean and smooth

$0.015/img and up View →
Nano Banana 2 Image

Nano Banana 2

nano-banana-2

High-quality text-to-image / image-to-image model

$0.025/img and up View →
Nano Banana Pro Image

Nano Banana Pro

nano-banana-pro

Flagship text-to-image / image-to-image model

$0.040/img and up View →
Claude Sonnet 4.6 Chat

Claude Sonnet 4.6

claude-sonnet-4-6

Balanced, efficient chat and coding model

$1.50/1M in · $7.50/1M out ↓50% View →
Claude Opus 5 Chat

Claude Opus 5

claude-opus-5

Next-generation flagship from Anthropic

$4.00/1M in · $20.00/1M out View →
Claude Fable 5 Chat

Claude Fable 5

claude-fable-5

Anthropic Fable series · narrative and long-form writing

$8.00/1M in · $40.00/1M out View →
Gemini 3.1 Pro Chat

Gemini 3.1 Pro

gemini-3.1-pro

Next-generation flagship reasoning model from Google

$0.50/1M in · $3.00/1M out ↓75% View →
Gemini 3.8 Flash Chat

Gemini 3.8 Flash

gemini-3.8-flash

Latest Flash · adaptive thinking

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.7 Flash Chat

Gemini 3.7 Flash

gemini-3.7-flash

Previous Flash · adaptive thinking

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.6 Flash Chat

Gemini 3.6 Flash

gemini-3.6-flash

Next-generation fast thinking model · four thinking budgets

$0.60/1M in · $3.60/1M out View →
Gemini 3 Flash Chat

Gemini 3 Flash

gemini-3-flash-preview

Fast, low-latency, cost-efficient chat model

$0.30/1M in · $1.20/1M out ↓57% View →
Gemini 3.5 Flash Chat

Gemini 3.5 Flash

gemini-3.5-flash

Next-generation fast thinking model

$0.45/1M in · $2.70/1M out ↓70% View →
Gemini 2.5 Flash Chat

Gemini 2.5 Flash

gemini-2.5-flash

Fast, low-latency, cost-efficient chat

$0.30/1M in · $1.20/1M out ↓46% View →
Veo 3.1 Video

Veo 3.1

veo-3.1

Text / image to video · multiple tiers · multiple resolutions

$0.075/clip+ View →
Seedance 2.5 Video

Seedance 2.5

seedance-2.5

Text / image to video · up to 30 seconds

$0.158/s View →
Wan 3.0 Video

Wan 3.0

wan3.0-video

One endpoint, five modes · up to 30 seconds

$0.120/s View →
MiniMax H3 Video

MiniMax H3

minimax-h3

1080P · native audio

$0.036/s View →
Grok Imagine Video 1.5 Video

Grok Imagine Video 1.5

grok-imagine-video-1.5

Billed per clip · up to 15 seconds

$0.300/clip+ View →

Why choose NezhaGate

Built for developers and AI applications — a real platform from day one.

Developer-first

OpenAI-compatible, zero rework

Just swap the Base URL and Key, and your existing SDK works as-is. Copy the Base URL, Bearer Token, and Python / Node.js examples with one click — or hand the whole page to an AI.

Operable

Full control over keys, quotas, and usage

A dedicated Key and limit for each project, with real-time balance and call logs. Admins can adjust pricing, top up, and enable or disable access.

Transparent

Pay-as-you-go, no charge on failure

Clear, transparent pricing settled by actual usage; upstream failures fall back automatically and are never billed to you.

FAQ

NezhaGate FAQ

Common questions on access, billing, models and safety.

What is NezhaGate?+
NezhaGate is an OpenAI-compatible AI model gateway that connects you to GPT, Claude, Gemini and image-generation models through one API and one key. Instead of integrating each provider separately, you point your Base URL at NezhaGate and call multiple flagship models with the OpenAI SDK you already use. The platform includes API key management, balance and usage tracking, a model market and developer docs, billed by usage with no charge on failed requests.
Which models are supported?+
NezhaGate currently provides the GPT-5.6 family (Sol / Terra / Luna), GPT-5.5 and GPT Image 2, Claude Sonnet 4.6, Gemini 3.1 Pro and Gemini 3.6 / 3.5 / 3 / 2.5 Flash, plus Nano Banana image generation and Veo 3.1 video, with new models added over time. The full, live model list and capability notes are shown on the Model Market page. Every model is called through the same endpoint and key.
How does NezhaGate bill?+
NezhaGate uses prepaid credits and pay-as-you-go billing: chat models are billed by input / output tokens and image models by count and resolution, with unit prices shown on the Pricing page and Model Market. There are no monthly plans — you pay for what you use, and the balance is drawn down by actual usage. Every call's usage and cost is visible in the console.
Am I charged for failed requests?+
Failed requests are not charged. Only calls that return successfully are billed by actual usage; upstream errors or timeouts do not consume your balance. If you see a failed record in the console, no cost is deducted for it.
How do I get started quickly with NezhaGate?+
Getting started takes three steps: register an account, create an API key in the console, then point your Base URL at NezhaGate and add the key. Because it is fully OpenAI-compatible, existing OpenAI SDK code needs almost no changes — just swap base_url and api_key. See the developer docs for details.
Are my account and API keys safe?+
An API key is shown in full exactly once, at creation; afterwards the console shows only its prefix, and you can disable or delete it at any time. Each key carries its own spend cap, requests-per-minute limit, model allowlist and IP allowlist, so a leaked key never reaches your other projects. Every call and every balance change is itemised, so your bill always reconciles with your usage. Request content is used to fulfil the call and bill it -- never to train models.