NezhaGate
Chat

Claude Opus 5

claude-opus-5

Claude Opus 5 是 Anthropic 新一代旗舰模型,默认开启自适应思考,长于复杂推理、代码工程与 Agent 编排,支持图片理解、工具调用与提示缓存。完全兼容 OpenAI 接口,把 model 改成 claude-opus-5 即可接入;也可用原生 Anthropic Messages 接口让 Claude Code 直连。

文本对话深度推理代码图片理解提示缓存

Live Test · Playground

Try out Claude Opus 5 right here (available after login)。

Input

Advanced
联网搜索 · Web search
Memory · multi-turn
思考过程 · Thinking

Conversation

Start a conversationType a message below to begin

About Claude Opus 5

Claude Opus 5 is Anthropic's next-generation flagship model. Adaptive thinking is on by default: the model decides how long to reason per request, answering simple questions directly and working through harder ones first, with the reasoning visible on the native API and in the playground. It is strong at complex reasoning, code engineering, agent orchestration and long-context documents, and natively supports image input, tool use and prompt caching. NezhaGate serves it two ways: the OpenAI-compatible endpoint (point base_url at NezhaGate, set model to claude-opus-5, reuse your OpenAI SDK unchanged), and the native Anthropic Messages API (point ANTHROPIC_BASE_URL at NezhaGate's /anthropic so Claude Code connects directly, preserving thinking, tool_use and prompt caching). Billing is pay-as-you-go and failed requests are not charged.

Use cases

Complex code engineering

Cross-file refactors, hard bug hunts, architecture review and migration plans - with tool use it can drive an agent through multi-step changes.

Deep reasoning and analysis

Multi-step logic, trade-off analysis, data interpretation and technology selection, with adaptive thinking spending more budget on the hard cases.

Long documents and codebases

Summarize, extract from and answer questions over contracts, papers, manuals and large repositories; prompt caching makes repeated questions much cheaper.

Agents and tool orchestration

Native tool_use supports function calling and multi-turn tool loops, which suits automated workflows and engineering assistants.

Mixed text and image tasks

Read charts, screenshots and design mockups and respond with analysis or code; both base64 and remote image URLs are accepted.

How to choose

Choose Claude Opus 5 when the task is genuinely hard and you want the model to think before it answers - complex reasoning, large code projects and agent orchestration are its home turf. For narrative, scripts or long-form writing, its sibling Claude Fable 5 has a finer touch; for cheaper everyday chat and coding, use Claude Sonnet 4.6; for million-token context, consider Gemini 3.1 Pro. All of them switch with a single model field on the same endpoint.

FAQ

What is Claude Opus 5?
Claude Opus 5 is Anthropic's next-generation flagship chat and coding model, with adaptive thinking on by default. It excels at complex reasoning, code engineering and agent orchestration, and supports image input, tool use and prompt caching. On NezhaGate you can call it through the OpenAI-compatible API or the native Anthropic Messages API.
Why is my temperature ignored?
Anthropic deprecated temperature / top_p / top_k on this generation, and sending them makes the upstream reject the call. So that existing OpenAI SDK code keeps working unchanged, the gateway strips those three parameters before forwarding: the request succeeds, no error is raised, and sampling runs at the model default. To steer the output, use reasoning_effort (low / medium / high / xhigh / max) to set thinking depth instead — all five rungs cost the same per token, deeper thinking simply emits more.
How do I see the thinking?
Call the native Anthropic Messages API and the response carries thinking content blocks; the NezhaGate playground shows the thinking summary directly. The OpenAI-compatible endpoint returns only the final answer, while thinking tokens are still billed as output.
Is prompt caching supported, and how is it billed?
Yes - cache writes and cache hits both verified working. Cache writes bill at the cache-write rate and hits at the cache-read rate, both listed on the pricing page. Long prompts may be cached upstream even without an explicit cache_control, which makes repeated calls noticeably cheaper.
Can Claude Code connect directly?
Yes. Point ANTHROPIC_BASE_URL at NezhaGate's /anthropic, use your gateway key as x-api-key and set model to claude-opus-5. Thinking, tool_use, prompt caching and streaming SSE all pass through unchanged.

Related Models

Explore other models you can integrate.

View all →
GPT-5.6 Sol Chat

GPT-5.6 Sol

gpt-5.6-sol

前沿旗舰 Agentic 编程模型

$2.00/1M tokens ↓75% View →
GPT-5.6 Terra Chat

GPT-5.6 Terra

gpt-5.6-terra

日常均衡 Agentic 编程模型

$1.20/1M tokens ↓77% View →
GPT-5.6 Luna Chat

GPT-5.6 Luna

gpt-5.6-luna

快而省的 Agentic 编程模型

$0.80/1M tokens ↓68% View →
GPT-5.5 Chat

GPT-5.5

gpt-5.5

旗舰对话与推理模型

$0.70/1M tokens ↓86% View →
GPT Image 2 Image

GPT Image 2

gpt-image-2

高质量文生图 / 图生图模型

$0.015/img and up View →
Nano Banana 2 Image

Nano Banana 2

nano-banana-2

高质量文生图 / 图生图模型

$0.013/img and up View →
Nano Banana Pro Image

Nano Banana Pro

nano-banana-pro

旗舰级文生图 / 图生图模型

$0.020/img and up View →
Claude Sonnet 4.6 Chat

Claude Sonnet 4.6

claude-sonnet-4-6

均衡高效的对话 / 代码模型

$1.50/1M tokens ↓50% View →
Claude Fable 5 Chat

Claude Fable 5

claude-fable-5

Anthropic Fable 系列 · 叙事与长文特化

$8.00/1M tokens View →
Gemini 3.1 Pro Chat

Gemini 3.1 Pro

gemini-3.1-pro

Google 新一代旗舰推理模型

$0.50/1M tokens ↓75% View →
Gemini 3.6 Flash Chat

Gemini 3.6 Flash

gemini-3.6-flash

新一代高速思考模型 · 四档思考预算

$0.60/1M tokens View →
Gemini 3.6 Flash High Chat

Gemini 3.6 Flash High

gemini-3.6-flash-high

深度思考档

$0.60/1M tokens View →
Gemini 3.6 Flash Low Chat

Gemini 3.6 Flash Low

gemini-3.6-flash-low

极速低耗档

$0.60/1M tokens View →
Gemini 3.6 Flash Tiered Chat

Gemini 3.6 Flash Tiered

gemini-3.6-flash-tiered

自动调档

$0.60/1M tokens View →
Gemini 3 Flash Chat

Gemini 3 Flash

gemini-3-flash-preview

高速低延迟,高性价比对话模型

$0.30/1M tokens ↓60% View →
Gemini 3.5 Flash Chat

Gemini 3.5 Flash

gemini-3.5-flash

新一代高速思考模型

$0.45/1M tokens ↓70% View →
Gemini 2.5 Flash Chat

Gemini 2.5 Flash

gemini-2.5-flash

高速低延迟,高性价比对话

$0.30/1M tokens ↓52% View →
Veo 3.1 Video

Veo 3.1

veo-3.1

文生 / 图生视频 · 多档 · 多分辨率

$0.075/clip+ View →
Seedance 2.5 Video

Seedance 2.5

seedance-2.5

文生 / 图生视频 · 最长 30 秒

$0.632/clip+ View →
Seedance 2.0 Video

Seedance 2.0

seedance-2.0

文生 / 图生视频 · 4–15 秒

$0.412/clip+ View →
Seedance 2.0 Fast Video

Seedance 2.0 Fast

seedance-2.0-fast

低延迟 · 最省

$0.332/clip+ View →
MiniMax H3 Video

MiniMax H3

minimax-h3

1440P · 原生音频

$0.360/clip+ View →