NezhaGate
Model Market

AI Model Market

Browse every available AI model, compare pricing and capabilities, and integrate in minutes. Currently offering 21 models, with more added continuously. Plus 1 not yet open.

GPT-5.6 Terra Chat

GPT-5.6

gpt-5.6 · sol · terra · luna

The GPT-5.6 family: Sol is the frontier flagship, Terra the balanced everyday model, Luna the fast and economical one -- a single family spanning frontier intelligence to high-speed execution. Use Model Type to pick a tier.

$1.20/1M in · $7.00/1M out ↓77% View →
GPT-5.5 Chat

GPT-5.5

gpt-5.5

GPT-5.5 delivers industry-leading chat, reasoning and coding, with streaming, long context and structured tasks. Fully OpenAI-SDK compatible: plug it into Codex, Cursor, Claude Code and similar tools with zero code changes.

$0.70/1M in · $4.20/1M out ↓86% View →
GPT-6 Astra Chat

GPT-6 Astra

gpt-6-astra

GPT-6 Astra is the next-generation flagship OpenAI released in September 2026: a 1,050,000-token context window, up to 128,000 output tokens, reasoning_effort up to xhigh, image input, tool calling and prompt caching. Fully OpenAI-SDK compatible - set model to gpt-6-astra - and served on both /chat/completions and /responses. Pay-as-you-go, failed calls never billed; cached input bills at one tenth, and we do not add the long-context surcharge the official API applies.

$2.80/1M in · $14.00/1M out ↓72% View →
GPT Image 2 Image

GPT Image 2

gpt-image-2

GPT Image 2 is built for commercial creative work: article covers, product shots, posters, infographics and illustrations. Supports both text-to-image and image-to-image, 1K / 2K / 4K resolutions and multiple aspect ratios, through the OpenAI-compatible images API.

$0.015/image and up View →
GPT Image 2.5 Flare Image

GPT Image 2.5

gpt-image-2.5 · flare · sunburst

GPT Image 2.5 is the next-generation OpenAI image model in two style branches: Flare comes back cleaner and smoother, Sunburst grainier with heavier texture and stronger highlights. Identical capabilities and price - only the look differs. Use Model Type to pick one.

$0.015/image and up View →
Nano Banana 2 Image

Nano Banana 2

nano-banana-2

Nano Banana 2 (Gemini 3.1 Flash Image) excels at covers, posters, illustrations and product shots. Supports text-to-image and image-to-image, 1K / 2K / 4K resolutions and multiple aspect ratios; fast and cost-efficient, through the OpenAI-compatible images API.

$0.025/image and up View →
Nano Banana Pro Image

Nano Banana Pro

nano-banana-pro

Nano Banana Pro (Gemini 3 Pro Image) is a flagship image model with stronger detail and composition. Supports text-to-image and image-to-image, excels at typography and complex scenes, and suits commercial covers and high-quality illustration, through the OpenAI-compatible images API.

$0.040/image and up View →
Claude Sonnet 4.6 Chat

Claude Sonnet 4.6

claude-sonnet-4-6

Claude Sonnet 4.6 strikes an excellent balance between capability and speed, excelling at everyday chat, code and agents at a great price. Fully OpenAI-API compatible: set model to claude-sonnet-4-6 and you are connected.

$1.50/1M in · $7.50/1M out ↓50% View →
Claude Opus 5 Chat

Claude Opus 5

claude-opus-5

Claude Opus 5 is the next-generation flagship from Anthropic, with adaptive thinking on by default. It excels at complex reasoning, code engineering and agent orchestration, and supports image understanding, tool use and prompt caching. Fully OpenAI-API compatible: set model to claude-opus-5; or use the native Anthropic Messages API so Claude Code connects directly.

$4.00/1M in · $20.00/1M out View →
Claude Fable 5 Chat

Claude Fable 5

claude-fable-5

Claude Fable 5 belongs to the Anthropic Fable series: same generation and capability base as Opus 5 (adaptive thinking, tool use, image understanding, prompt caching), with a finer touch and richer emotional layering in Chinese narrative and long-form writing. Fully OpenAI-API compatible: set model to claude-fable-5; the native Anthropic Messages API is supported as well.

$8.00/1M in · $40.00/1M out View →
Gemini 3.1 Pro Chat

Gemini 3.1 Pro

gemini-3.1-pro

Gemini 3.1 Pro is the next-generation flagship multimodal model from Google, with a very long context window and top-tier reasoning and coding. Fully OpenAI-API compatible: set model to gemini-3.1-pro in your existing SDK. Low-reasoning tier by default (faster and cheaper); pass reasoning_effort=high for the high tier (billed at the Preview rate). Responses carry a reasoning_content field so you can see the thinking (streamed chunk by chunk as well).

$0.50/1M in · $3.00/1M out ↓75% View →
Gemini 3.8 Flash Chat

Gemini 3.8 Flash

gemini-3.8-flash

Gemini 3.8 Flash is the latest Flash generation from Google (released September 2026), built for long-running coding, agent workflows and complex reasoning: 1M-token context, up to 64K output, image input. It uses a dynamic thinking budget -- quick answers for simple questions, more thinking for hard ones. Fully OpenAI-API compatible: set model to gemini-3.8-flash.

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.7 Flash Chat

Gemini 3.7 Flash

gemini-3.7-flash

Gemini 3.7 Flash is the generation before 3.8, with the same dynamic thinking budget, 1M-token context and image input. A good fit if your prompts are already tuned for 3.7 and you do not want to switch yet. Fully OpenAI-API compatible: set model to gemini-3.7-flash.

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.6 Flash Chat

Gemini 3.6 Flash

gemini-3.6-flash · high · low · tiered

The Gemini 3.6 Flash family: one model, four thinking budgets -- Medium for everyday balance, High for deep reasoning, Low for speed and low cost, Tiered lets the model pick by difficulty. All four cost the same; more thinking simply means more output tokens. 1M context, image input. Use Model Type to pick a tier.

$0.60/1M in · $3.60/1M out View →
Gemini 3 Flash Chat

Gemini 3 Flash

gemini-3-flash-preview

Gemini 3 Flash focuses on speed and cost efficiency, for high-volume, low-latency chat and agent workloads. Fully OpenAI-API compatible with streaming: set model to gemini-3-flash-preview.

$0.30/1M in · $1.20/1M out ↓57% View →
Gemini 3.5 Flash Chat

Gemini 3.5 Flash

gemini-3.5-flash

Gemini 3.5 Flash is the next-generation fast model from Google with built-in thinking: fast and cost-efficient, for high-volume chat and agent workloads. Fully OpenAI-API compatible: set model to gemini-3.5-flash.

$0.45/1M in · $2.70/1M out ↓70% View →
Gemini 2.5 Flash Chat

Gemini 2.5 Flash

gemini-2.5-flash

Gemini 2.5 Flash focuses on speed and cost efficiency, for large-scale, low-latency chat and agent workloads, with multimodal input. Fully OpenAI-API compatible: set model to gemini-2.5-flash.

$0.30/1M in · $1.20/1M out ↓46% View →
Veo 3.1 Video

Veo 3.1

veo-3.1

Veo 3.1 is the AI video model from Google. One endpoint, parameters select the mode: text-to-video without a reference image, image-to-video with one (used as the first frame). Quality tiers Lite / Fast / Quality, resolutions 720p / 1080p / 4K (1080p and 4K on the Quality tier), landscape 16:9 or portrait 9:16, durations 4s / 6s / 8s. Asynchronous: submit, then poll the job id for a stable re-hosted video link. Billed per clip; failed jobs are refunded in full.

$0.075/clip+ View →
Seedance 2.5 Video

Seedance 2.x

seedance-2.5

The ByteDance Seedance video family: 2.5 goes up to 30 seconds with the highest reference-input limits; 2.0 and 2.0 Fast are quicker and cheaper.

$0.158/s View →
Wan 3.0 Video

Wan 3.0

wan3.0-video

The Alibaba Wan 3.0 video family: Standard and Prime have identical capabilities; they differ only in turnaround time and unit price.

$0.120/s View →
MiniMax H3 Video

MiniMax H3

minimax-h3

MiniMax H3 (Hailuo 03) video model: 1080P HD output, native audio always generated, 5-15 seconds in one-second steps, six aspect ratios. Text-to-video without a reference image, image-to-video with one (used as the first frame). Asynchronous: submit, then poll the job id for a stable re-hosted mp4 link. Billed per second; failed jobs are refunded in full.

$0.036/s View →
Grok Imagine Video 1.5 Video

Grok Imagine Video 1.5

grok-imagine-video-1.5

xAI Grok Imagine Video 1.5: 4-15 seconds in one-second steps, 480P / 720P, five aspect ratios. Text-to-video without a reference image, image-to-video with reference images (up to 7; with 2 or more the maximum is 10 seconds). Asynchronous, billed per clip -- 1 second costs the same as 15 -- and failed jobs are refunded in full. 480P $0.30 / 720P $0.40; the price only changes with resolution.

$0.300/clip+ View →