Key Highlights
- Opus-level capability at Sonnet pricing: Anthropic reports 70.6% on Terminal-Bench 4.0, above Opus 5.5’s 66.4%; GDPval-AA knowledge work scores 1844, essentially tied with Opus 5.5’s 1846
- Unchanged pricing: $2 input / $10 output per million tokens and $0.20 cache reads, identical to Sonnet 5 and half the price of Opus 5.5
- Faster and more token-efficient: Anthropic says output is more than 30% faster than Sonnet 5 and the same work takes fewer tokens, cutting per-task cost by up to about 30%
- Live on APIYI:
claude-sonnet-5-5andclaude-sonnet-5-5-thinking, over both OpenAI-compatible and Anthropic-native endpoints, with all four billing items matching the official price - Five breaking changes:
thinking: disabledreturns 400 (usebetween_tools), forced tool use returns 400, thinking blocks are bound to the model and conversation, computer use only accepts the new toolset, and advisor pairings are narrower.top_p/top_kare no longer supported and return an error in APIYI testing
Background
On September 28, 2026, Anthropic released Claude Sonnet 5.5, the second model in the Claude 5.5 family after Opus 5.5 a week earlier. Anthropic also said Haiku 5.5 will follow in the coming weeks. The Sonnet line has always been the everyday workhorse: fast and moderately priced. Sonnet 5 brought it close to Opus 4.8 three months ago. Sonnet 5.5 is built on the same foundation as Opus 5.5 and now beats the flagship on agentic coding, trails it by only 2 points on knowledge work, and costs half as much. Anthropic describes its strongest areas as well-scoped everyday tasks, bug fixing, and producing polished documents, slides, and spreadsheets. It is also the first Sonnet model to beat Pokémon Red working only from screenshots, which Anthropic cites as evidence of its long-horizon and image-understanding abilities. APIYI has launchedclaude-sonnet-5-5 (plus claude-sonnet-5-5-thinking), and input, output, cache write, and cache read billing all match the official price.
Detailed Analysis
Core Features
Beats Opus 5.5 at Agentic Coding
70.6% on Terminal-Bench 4.0 (Anthropic), above Opus 5.5’s 66.4%; CursorBench 4.0 rises from Sonnet 5’s 34.1% to 55.5%
Faster, Fewer Tokens
More than 30% faster output than Sonnet 5; early testers report it groups tool calls into fewer steps
1M Context
1 million token context, 128K max output, and the same tokenizer as Sonnet 5, so the same text gives the same token count
Flagship-Level Knowledge Work
GDPval-AA 1844 and AA-Briefcase 1811, nearly tied with Opus 5.5 (1846 / 1822) and about 400 points above Sonnet 5
Benchmarks (Vendor-Reported)
Source: Anthropic’s announcement page
anthropic.com/claude-sonnet-5-5 (September 28, 2026). Anthropic’s footnotes: the Opus 5.5 Terminal-Bench score is at xhigh effort and FrontierCode at max effort; some GPT-6 Sol scores are outdated because of a recent bug fix on that side. Retrieved September 30, 2026.Independent Evaluation (Artificial Analysis)
On the same benchmark, the independent evaluation reports lower scores than Anthropic, but the ranking is the same:For Developers: Breaking Changes
Sonnet 5.5 keeps the Sonnet 5 request structure, but the following changes return errors or change the response shape. Check them before migrating:
Two more changes do not cause errors:
- Progress text between tool calls now arrives in
thinkingblocks: with the defaultdisplay: "omitted"their text is empty, so an interface that streams that text to users goes quiet between tool calls. Usebetween_toolsor setthinking.displayto get it back - Safety classifiers: a declined request returns HTTP 200 with
stop_reason: "refusal", andstop_detailsnames one of five categories (cyber/bio/frontier_llm/reasoning_extraction/general_harms). This is the first Sonnet with cyber and anti-distillation classifiers; higher-risk cyber requests fall back to Sonnet 5 on the provider side
Technical Specifications
APIYI Test Results
Tested on APIYI on September 30, 2026 (UTC+8):
The last two rows show where cost comes from: the same model spends almost no thinking tokens on a simple question, and nearly ten thousand on a hard one at
max. The effort level affects your bill far more than the unit price does.
Practical Applications
Recommended Use Cases
- Everyday coding: bug fixes, small repository changes, and terminal agent tasks; Terminal-Bench 4.0 shows the biggest gain
- Documents, slides, and spreadsheets: an area Anthropic highlights, with AA-Briefcase and GDPval-AA close to Opus 5.5
- High-frequency agents and support workflows: faster output and fewer tool-call steps suit latency-sensitive multi-turn tool use
- Replacing Sonnet 5 and some Opus 5.5 workloads: a same-price upgrade from Sonnet 5; well-scoped tasks on Opus 5.5 are worth evaluating on Sonnet 5.5 at about half the cost
Code Examples
Anthropic-Native Format
Turning Off Up-Front Thinking (Replaces disabled)
OpenAI-Compatible Format
Six-Step Migration Checklist from Sonnet 5
- Change the model name:
claude-sonnet-5→claude-sonnet-5-5 - Replace the thinking-off setting: change
thinking: {"type": "disabled"}tobetween_tools, with effort no higher thanhigh - Remove forced tool use: set
tool_choicetoautoand usestrict: trueto keep tool input valid - Drop sampling parameters:
top_p/top_kreturn an error andtemperaturehas no effect, so send none of them; stop using manualbudget_tokenstoo and control thinking depth witheffort - Recalibrate effort: the same level does not produce the same amount of thinking as on Sonnet 5, so re-run your evaluation instead of carrying the old setting over
- Keep history append-only: pass thinking blocks back unchanged and do not edit
system/tools/ earlier messages between requests
Pricing and Availability
Pricing
Official prices (USD per million tokens):
The Sonnet 5.5 column is also APIYI’s price: all four billing items match the official price, with no markup.
Model prices are aligned with the official website and may change with it; the table above is for reference only, and the Model Pricing tab in the top navigation is authoritative: Model Pricing.
Our pricing matches the official price, with no markup on the model. On top of that, some groups carry an additional discount, and recharge bonuses can be stacked. That is our own discount and separate from the model’s pricing.For actual charges, rely on the live data in the console’s cache billing details. The cache fields in the API’s usage response are not a billing reference.
Groups and Endpoints
The
ClaudeCode group is for Anthropic-native clients such as Claude Code. Match the protocol to the group: use the ClaudeCode group for the Anthropic-native protocol and the default / svip groups for the OpenAI-compatible protocol. The ClaudeCode group carries an additional discount that stacks with recharge bonuses.Stack with Recharge Promotions
APIYI recharge bonuses can lower your effective cost further. See Recharge Promotions.Summary and Recommendations
Claude Sonnet 5.5 is a same-price upgrade: it costs exactly what Sonnet 5 did, beats Opus 5.5 on agentic coding, trails the flagship by only 2 points on knowledge work, and runs about 30% faster. For most everyday coding and document work, it is currently the best value in the Claude lineup. Recommendations:- Sonnet 5 users: upgrade directly, but go through the six-step checklist first, especially
disabled→between_tools, forced tool use, and sampling parameters - Opus 5.5 users: evaluate well-scoped coding and document tasks on Sonnet 5.5 first at about half the cost; keep complex architecture decisions and long research chains on Opus 5.5
- Cost control: do not default to
maxeffort; start atmedium/highand raise it only if results call for it
Sources: Anthropic’s announcement page
anthropic.com/claude-sonnet-5-5; Anthropic developer docs platform.claude.com/docs/en/models/sonnet-5-5/overview and whats-new-sonnet-5-5; Artificial Analysis independent evaluation (via Decrypt); reporting from SiliconANGLE, Unite.AI, Help Net Security, and others. The APIYI Test Results section comes from our own calls on September 30, 2026. APIYI pricing follows live platform data. Retrieved September 30, 2026.