Doubao Seed 2.1 Turbo
⧉Doubao Seed 2.1 Turbo is the lower-cost member of the Doubao Seed 2.1 family, at about half the unit price of Pro, for everyday chat and high-volume jobs. By default it thinks before it answers, with the reasoning in reasoning_content billed at the output rate (reasoning_effort: "minimal" usually skips it, though less reliably than on Pro); tool calling, JSON output, streaming and image input (base64 or public links) are supported. A repeated long prefix can hit the cache automatically, and that part settles at the cache rate. Fully OpenAI-compatible: set model to doubao-seed-2-1-turbo. Pay-as-you-go, failed calls never billed.
Live Test · Playground
Try out Doubao Seed 2.1 Turbo right here (available after login).
Input
Advanced
Conversation
About Doubao Seed 2.1 Turbo
Doubao Seed 2.1 Turbo is the lower-cost member of the Doubao Seed 2.1 family, served on NezhaGate through the OpenAI-compatible API. It costs about half as much per token as Pro, which suits everyday chat and high-volume jobs, and it accepts image input. By default it thinks before it answers: the reasoning comes back in message.reasoning_content and the answer in content, and reasoning tokens bill at the output rate. Tool calling, JSON output, streaming and prompt caching are supported. Set model to doubao-seed-2-1-turbo; pay-as-you-go, failed calls never billed.
Use cases
A low unit price for high-frequency chat such as support, Q&A and everyday writing; switching models is just the model field.
Classification, extraction, summarising and rewriting in bulk, at about half the unit price of Pro.
Q&A over screenshots, receipts and charts; in our tests it read the text in an image accurately. Send images as base64 or public links.
A repeated long prefix hits the cache automatically (9,528 of 9,828 input tokens on the second call in our tests), and the cached part settles at the cache rate.
How to choose
Choose Turbo for everyday chat, high-volume jobs or when the unit price matters most, with automatic caching of a repeated long prefix; choose Pro for the flagship reasoning, coding and agent work of the Doubao Seed 2.1 series. Both belong to the Doubao Seed 2.1 family, so switching is just the model field.
FAQ
Can I turn thinking off?
Can max_tokens limit the output?
Is prompt caching supported?
How do I choose between Turbo and Pro?
Am I charged for failed calls?
Related Models
Explore other models you can integrate.
Doubao Seed 2.1 Pro
ByteDance's Doubao flagship - reasoning, coding and agents
Qwen3.8 Flash
The light, fast Qwen3.8 - high volume, low cost
DeepSeek V4.1 Flash
DeepSeek's current Flash model - thinking on or off
GLM-5.3 Flash
The light GLM-5.3 - high volume, low cost
Gemini 3.6 Flash
Next-generation fast thinking model · four thinking budgets