Claude Sonnet 5
⧉Claude Sonnet 5 is Anthropic's new-generation Sonnet: Sonnet-tier speed and price, with coding and agentic work close to Opus. Adaptive thinking is on by default, and five effort levels (low to max) trade depth for speed. It supports image understanding, tool use and prompt caching. Fully OpenAI-API compatible: set model to claude-sonnet-5; or use the native Anthropic Messages API so Claude Code connects directly.
Live Test · Playground
Try out Claude Sonnet 5 right here (available after login).
Input
Advanced
Conversation
About Claude Sonnet 5
Claude Sonnet 5 is the newest model in Anthropic's Sonnet line - the Sonnet of the Claude 5 generation. It keeps Sonnet-tier speed and price while bringing coding and agentic work close to Opus quality. Adaptive thinking is on by default: the model decides how long to reason, answering simple requests directly and working through harder ones first, and five effort levels (low / medium / high / xhigh / max) let you trade depth for speed. It supports image input, tool use and prompt caching. NezhaGate serves it two ways: the OpenAI-compatible endpoint (point base_url at NezhaGate, set model to claude-sonnet-5, reuse your OpenAI SDK unchanged), and the native Anthropic Messages API (point ANTHROPIC_BASE_URL at NezhaGate's /anthropic so Claude Code connects directly, with thinking, tool_use and prompt caching intact). Billing is pay-as-you-go and failed requests are not charged.
Use cases
Write, refactor and debug code, and run multi-step agent loops with tool calls - a strong default for coding assistants and bots that make many calls a day.
Support replies, extraction, summarization and classification at scale, where you want quality close to Opus without paying Opus rates.
Summarize, extract from and answer questions over contracts, papers, manuals and repositories; prompt caching makes repeated questions over the same material much cheaper.
Keep effort at low for quick chat and raise it to high or xhigh for hard bugs and planning; the per-token price is the same at every level.
Read screenshots, charts and design mockups and respond with analysis or code; both base64 and remote image URLs are accepted.
How to choose
Make Claude Sonnet 5 your default Claude for most work: it is faster and cheaper than Claude Opus 5 and close to it on coding and agentic tasks. Step up to Claude Opus 5 for the hardest reasoning and the largest code projects, or switch to Claude Fable 5 for narrative and long-form writing; keep Claude Sonnet 4.6 for workflows already tuned to it. All of them switch with a single model field on the same endpoint.
FAQ
What is Claude Sonnet 5?
How is it different from Claude Sonnet 4.6?
Why is my temperature ignored?
Can I turn thinking off?
Can Claude Code connect directly?
Related Models
Explore other models you can integrate.
Claude Opus 5
Next-generation flagship from Anthropic
Claude Sonnet 4.6
Balanced, efficient chat and coding model
Claude Fable 5
Anthropic Fable series · narrative and long-form writing
GPT-6 Sol
Built for complex coding and agentic workflows - 1.05M context
Gemini 3.1 Pro
Next-generation flagship reasoning model from Google
Grok 4.7
xAI's flagship reasoning model - long context