🚀 New · GPT-5.6 family is live: Sol (frontier), Terra (balanced), Luna (fast & affordable) — ready on the OpenAI-compatible API. Explore →
NezhaGate
OpenAI-Compatible AI Gateway

One API,Three flagships

NezhaGate / NezhaAPI is an OpenAI-compatible AI model gateway that helps developers connect to GPT, Claude, Gemini and image-generation models through one API. Built-in API key management, balance and usage, a model market and developer docs — billed by usage, with no charge on failed requests.

🎁 New users get 80 free credits — Google / GitHub login supported. Sign up free →

OpenAI SDK compatibleStreaming outputPay-as-you-go billingNo charge on failure
99.9%Availability SLA
OpenAIAPI format compatible
Meteredbilling · no hidden fees
17open models · always growing

Model market

Top models hand-picked across every category — start integrating in minutes.

View all →
GPT-5.6 Sol Chat

GPT-5.6 Sol

gpt-5.6-sol

前沿旗舰 Agentic 编程模型

$2.00/1M tokens ↓75% View →
GPT-5.6 Terra Chat

GPT-5.6 Terra

gpt-5.6-terra

日常均衡 Agentic 编程模型

$1.20/1M tokens ↓77% View →
GPT-5.6 Luna Chat

GPT-5.6 Luna

gpt-5.6-luna

快而省的 Agentic 编程模型

$0.80/1M tokens ↓68% View →
GPT-5.5 Chat

GPT-5.5

gpt-5.5

旗舰对话与推理模型

$0.70/1M tokens ↓86% View →
GPT Image 2 Image

GPT Image 2

gpt-image-2

高质量文生图 / 图生图模型

$0.015/img and up View →
Nano Banana 2 Image

Nano Banana 2

nano-banana-2

高质量文生图 / 图生图模型

$0.025/img and up View →
Nano Banana Pro Image

Nano Banana Pro

nano-banana-pro

旗舰级文生图 / 图生图模型

$0.040/img and up View →
Claude Sonnet 4.6 Chat

Claude Sonnet 4.6

claude-sonnet-4-6

均衡高效的对话 / 代码模型

$1.50/1M tokens ↓50% View →
Gemini 3.1 Pro Chat

Gemini 3.1 Pro

gemini-3.1-pro

Google 新一代旗舰推理模型

$0.50/1M tokens ↓75% View →
Gemini 3.6 Flash Chat

Gemini 3.6 Flash

gemini-3.6-flash

新一代高速思考模型 · 四档思考预算

$0.60/1M tokens View →
Gemini 3.6 Flash High Chat

Gemini 3.6 Flash High

gemini-3.6-flash-high

深度思考档

$0.60/1M tokens View →
Gemini 3.6 Flash Low Chat

Gemini 3.6 Flash Low

gemini-3.6-flash-low

极速低耗档

$0.60/1M tokens View →
Gemini 3.6 Flash Tiered Chat

Gemini 3.6 Flash Tiered

gemini-3.6-flash-tiered

自动调档

$0.60/1M tokens View →
Gemini 3 Flash Chat

Gemini 3 Flash

gemini-3-flash-preview

高速低延迟,高性价比对话模型

$0.30/1M tokens ↓60% View →
Gemini 3.5 Flash Chat

Gemini 3.5 Flash

gemini-3.5-flash

新一代高速思考模型

$0.45/1M tokens ↓70% View →
Gemini 2.5 Flash Chat

Gemini 2.5 Flash

gemini-2.5-flash

高速低延迟,高性价比对话

$0.30/1M tokens ↓52% View →
Veo 3.1 Video

Veo 3.1

veo-3.1

文生 / 图生视频 · 多档 · 多分辨率

$0.075/clip+ View →

Why choose NezhaGate

Built for developers and AI applications — a real platform from day one.

Developer-first

OpenAI-compatible, zero rework

Just swap the Base URL and Key, and your existing SDK works as-is. Copy the Base URL, Bearer Token, and Python / Node.js examples with one click — or hand the whole page to an AI.

Operable

Full control over keys, quotas, and usage

A dedicated Key and limit for each project, with real-time balance and call logs. Admins can adjust pricing, top up, and enable or disable access.

Transparent

Pay-as-you-go, no charge on failure

Clear, transparent pricing settled by actual usage; upstream failures fall back automatically and are never billed to you.

FAQ

NezhaGate FAQ

Common questions on access, billing, models and safety.

What is NezhaGate?+
NezhaGate is an OpenAI-compatible AI model gateway that connects you to GPT, Claude, Gemini and image-generation models through one API and one key. Instead of integrating each provider separately, you point your Base URL at NezhaGate and call multiple flagship models with the OpenAI SDK you already use. The platform includes API key management, balance and usage tracking, a model market and developer docs, billed by usage with no charge on failed requests.
Which models are supported?+
NezhaGate currently provides the GPT-5.6 family (Sol / Terra / Luna), GPT-5.5 and GPT Image 2, Claude Sonnet 4.6, Gemini 3.1 Pro and Gemini 3.6 / 3.5 / 3 / 2.5 Flash, plus Nano Banana image generation and Veo 3.1 video, with new models added over time. The full, live model list and capability notes are shown on the Model Market page. Every model is called through the same endpoint and key.
How does NezhaGate bill?+
NezhaGate uses prepaid credits and pay-as-you-go billing: chat models are billed by input / output tokens and image models by count and resolution, with unit prices shown on the Pricing page and Model Market. There are no monthly plans — you pay for what you use, and the balance is drawn down by actual usage. Every call's usage and cost is visible in the console.
Am I charged for failed requests?+
Failed requests are not charged. Only calls that return successfully are billed by actual usage; upstream errors or timeouts do not consume your balance. If you see a failed record in the console, no cost is deducted for it.
How do I get started quickly with NezhaGate?+
Getting started takes three steps: register an account, create an API key in the console, then point your Base URL at NezhaGate and add the key. Because it is fully OpenAI-compatible, existing OpenAI SDK code needs almost no changes — just swap base_url and api_key. See the developer docs for details.
Is there a ban risk?+
Under compliant use, NezhaGate aims to provide stable, continuous model access. Actual call behavior can be affected by the model, route, request volume and upstream status. You can review your API keys, balance and every call record in the console.