NezhaGate
Video

Grok Imagine Video 1.5

grok-imagine-video-1.5

xAI Grok Imagine Video 1.5: 4-15 seconds in one-second steps, 480P / 720P, five aspect ratios. Text-to-video without a reference image, image-to-video with reference images (up to 7; with 2 or more the maximum is 10 seconds). Asynchronous, billed per clip -- 1 second costs the same as 15 -- and failed jobs are refunded in full. 480P $0.30 / 720P $0.40; the price only changes with resolution.

Text-to-videoImage-to-videoPer-clip pricingAsync jobs

Live Test · Playground

Try out Grok Imagine Video 1.5 right here (available after login)。

Input

Video
0 / 5000 bytes
With multiple references, address them in the prompt as @Image1 / @Image2, e.g. "the character from @Image1 stands in the scene from @Image2". Without tokens the model decides on its own.
Rendering takes ~1-3 min; billed per clip (any length, same price), fully refunded on failure. Prompt limit 5000 bytes (about 1666 Chinese characters).

Output

video
Results will appear here after running.

About Grok Imagine Video 1.5

Grok Imagine Video 1.5 is xAI's video model and the cheapest video model on NezhaGate. What sets it apart from everything else here is the billing shape: it is priced PER CLIP, so a 1-second and a 15-second render cost exactly the same; only the resolution changes the price (480P $0.30, 720P $0.40). Duration is any integer 1-15 seconds, resolution 480P / 720P, aspect ratio 16:9 / 9:16 / 1:1 / 4:3 / 3:4; omit the reference image for text-to-video, include up to 7 for image-to-video; every clip ships with a native audio track. Async job: POST /v1/videos/generations returns HTTP 202 and a job id immediately; poll GET /v1/videos/jobs/{id} until status=succeeded and read data[0].url. That link is an mp4 re-hosted by us -- it will not expire the way the upstream CDN link does.

Use cases

Batch idea validation

The lowest fixed cost per clip on the platform, so you can run dozens of prompts and keep the best one.

Long clips are the bargain

Per-clip pricing means the 15-second option is the best value; there is no reason to cut a clip short to save money.

Vertical social assets

9:16 straight out of the model for short-video platforms; 4:3 and 3:4 suit mixed image-and-text layouts.

Multi-image references

Up to 7 reference images, useful for putting one character or product into different scenes.

How to choose

Pick it when budget comes first or when you are batching ideas -- it starts at $0.30, the cheapest video here, and per-clip pricing means 15 seconds costs no more than one. For a controllable soundtrack, reference audio or reference video use the Seedance family; for more than 15 seconds use seedance-2.5; for 1080P use minimax-h3.

FAQ

Why per clip instead of per second?
Because the upstream charges per call: the cost is identical whatever length you ask for. We match the shape of our sell price to the shape of our cost rather than copying a per-second number that would overcharge short clips and lose money on long ones. Practically: within your chosen resolution, just ask for the longest duration you can use. The two resolutions cost us the same upstream; the split exists so that people who want the cheaper option have one.
Why only 10 seconds with multiple images?
That is an upstream limit: 15 seconds with a single reference image, 10 seconds with 2-7. We reject it at submit time with a clear message rather than letting you wait several minutes for a failure.
Does it produce audio?
Yes -- every measured render carries a native AAC stereo track. But the upstream endpoint has no switch and accepts no reference audio, so you can neither turn it off nor supply your own. For a controllable soundtrack use the Seedance family (audio parameter).
How is it billed, and am I charged on failure?
Per clip: the hold placed at submit is one clip's price and the settle is the exact same number. A failed or timed-out render refunds the whole hold, so you never pay for a clip you did not receive.

Related Models

Explore other models you can integrate.

View all →
GPT-5.6 Sol Chat

GPT-5.6 Sol

gpt-5.6-sol

Frontier flagship for agentic coding

$2.00/1M in · $12.00/1M out ↓75% View →
GPT-5.6 Terra Chat

GPT-5.6 Terra

gpt-5.6-terra

Balanced everyday agentic coding model

$1.20/1M in · $7.00/1M out ↓77% View →
GPT-5.6 Luna Chat

GPT-5.6 Luna

gpt-5.6-luna

Fast, economical agentic coding model

$0.80/1M in · $4.80/1M out ↓68% View →
GPT-5.5 Chat

GPT-5.5

gpt-5.5

Flagship chat and reasoning model

$0.70/1M in · $4.20/1M out ↓86% View →
GPT-6 Astra Chat

GPT-6 Astra

gpt-6-astra

Next-generation OpenAI flagship - 1.05M context

$2.80/1M in · $14.00/1M out ↓72% View →
GPT Image 2 Image

GPT Image 2

gpt-image-2

High-quality text-to-image / image-to-image model

$0.015/img and up View →
GPT Image 2.5 Flare Image

GPT Image 2.5 Flare

gpt-image-2.5-flare

Next-gen image model - clean and smooth

$0.015/img and up View →
GPT Image 2.5 Sunburst Image

GPT Image 2.5 Sunburst

gpt-image-2.5-sunburst

Next-gen image model - richer texture

$0.015/img and up View →
Nano Banana 2 Image

Nano Banana 2

nano-banana-2

High-quality text-to-image / image-to-image model

$0.025/img and up View →
Nano Banana Pro Image

Nano Banana Pro

nano-banana-pro

Flagship text-to-image / image-to-image model

$0.040/img and up View →
Claude Sonnet 4.6 Chat

Claude Sonnet 4.6

claude-sonnet-4-6

Balanced, efficient chat and coding model

$1.50/1M in · $7.50/1M out ↓50% View →
Claude Opus 5 Chat

Claude Opus 5

claude-opus-5

Next-generation flagship from Anthropic

$4.00/1M in · $20.00/1M out View →
Claude Fable 5 Chat

Claude Fable 5

claude-fable-5

Anthropic Fable series · narrative and long-form writing

$8.00/1M in · $40.00/1M out View →
Gemini 3.1 Pro Chat

Gemini 3.1 Pro

gemini-3.1-pro

Next-generation flagship reasoning model from Google

$0.50/1M in · $3.00/1M out ↓75% View →
Gemini 3.8 Flash Chat

Gemini 3.8 Flash

gemini-3.8-flash

Latest Flash · adaptive thinking

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.7 Flash Chat

Gemini 3.7 Flash

gemini-3.7-flash

Previous Flash · adaptive thinking

$0.60/1M in · $3.60/1M out ↓7% View →
Gemini 3.6 Flash Chat

Gemini 3.6 Flash

gemini-3.6-flash

Next-generation fast thinking model · four thinking budgets

$0.60/1M in · $3.60/1M out View →
Gemini 3.6 Flash High Chat

Gemini 3.6 Flash High

gemini-3.6-flash-high

Deep thinking tier

$0.60/1M in · $3.60/1M out View →
Gemini 3.6 Flash Low Chat

Gemini 3.6 Flash Low

gemini-3.6-flash-low

Fastest, lowest-cost tier

$0.60/1M in · $3.60/1M out View →
Gemini 3.6 Flash Tiered Chat

Gemini 3.6 Flash Tiered

gemini-3.6-flash-tiered

Auto-tiered thinking

$0.60/1M in · $3.60/1M out View →
Gemini 3 Flash Chat

Gemini 3 Flash

gemini-3-flash-preview

Fast, low-latency, cost-efficient chat model

$0.30/1M in · $1.20/1M out ↓57% View →
Gemini 3.5 Flash Chat

Gemini 3.5 Flash

gemini-3.5-flash

Next-generation fast thinking model

$0.45/1M in · $2.70/1M out ↓70% View →
Gemini 2.5 Flash Chat

Gemini 2.5 Flash

gemini-2.5-flash

Fast, low-latency, cost-efficient chat

$0.30/1M in · $1.20/1M out ↓46% View →
Veo 3.1 Video

Veo 3.1

veo-3.1

Text / image to video · multiple tiers · multiple resolutions

$0.075/clip+ View →
Seedance 2.5 Video

Seedance 2.5

seedance-2.5

Text / image to video · up to 30 seconds

$0.158/s View →
Seedance 2.0 Video

Seedance 2.0

seedance-2.0

Text / image to video · 5-15 seconds

$0.150/s View →
Seedance 2.0 Fast Video

Seedance 2.0 Fast

seedance-2.0-fast

Low latency · lowest cost

$0.100/s View →
Seedance 2.0 Mini Video

Seedance 2.0 Mini

seedance-2.0-mini

Entry tier · lowest cost

$0.066/s View →
Wan 3.0 Video

Wan 3.0

wan3.0-video

One endpoint, five modes · up to 30 seconds

$0.120/s View →
Wan 3.0 Prime Video

Wan 3.0 Prime

wan3.0-video-prime

Same capabilities · several times faster

$0.160/s View →
MiniMax H3 Video

MiniMax H3

minimax-h3

1080P · native audio

$0.036/s View →