Skip to main content

Key Highlights

  • Opus-level capability at Sonnet pricing: Anthropic reports 70.6% on Terminal-Bench 4.0, above Opus 5.5’s 66.4%; GDPval-AA knowledge work scores 1844, essentially tied with Opus 5.5’s 1846
  • Unchanged pricing: $2 input / $10 output per million tokens and $0.20 cache reads, identical to Sonnet 5 and half the price of Opus 5.5
  • Faster and more token-efficient: Anthropic says output is more than 30% faster than Sonnet 5 and the same work takes fewer tokens, cutting per-task cost by up to about 30%
  • Live on APIYI: claude-sonnet-5-5 and claude-sonnet-5-5-thinking, over both OpenAI-compatible and Anthropic-native endpoints, with all four billing items matching the official price
  • Five breaking changes: thinking: disabled returns 400 (use between_tools), forced tool use returns 400, thinking blocks are bound to the model and conversation, computer use only accepts the new toolset, and advisor pairings are narrower. top_p / top_k are no longer supported and return an error in APIYI testing

Background

On September 28, 2026, Anthropic released Claude Sonnet 5.5, the second model in the Claude 5.5 family after Opus 5.5 a week earlier. Anthropic also said Haiku 5.5 will follow in the coming weeks. The Sonnet line has always been the everyday workhorse: fast and moderately priced. Sonnet 5 brought it close to Opus 4.8 three months ago. Sonnet 5.5 is built on the same foundation as Opus 5.5 and now beats the flagship on agentic coding, trails it by only 2 points on knowledge work, and costs half as much. Anthropic describes its strongest areas as well-scoped everyday tasks, bug fixing, and producing polished documents, slides, and spreadsheets. It is also the first Sonnet model to beat Pokémon Red working only from screenshots, which Anthropic cites as evidence of its long-horizon and image-understanding abilities. APIYI has launched claude-sonnet-5-5 (plus claude-sonnet-5-5-thinking), and input, output, cache write, and cache read billing all match the official price.

Detailed Analysis

Core Features

Beats Opus 5.5 at Agentic Coding

70.6% on Terminal-Bench 4.0 (Anthropic), above Opus 5.5’s 66.4%; CursorBench 4.0 rises from Sonnet 5’s 34.1% to 55.5%

Faster, Fewer Tokens

More than 30% faster output than Sonnet 5; early testers report it groups tool calls into fewer steps

1M Context

1 million token context, 128K max output, and the same tokenizer as Sonnet 5, so the same text gives the same token count

Flagship-Level Knowledge Work

GDPval-AA 1844 and AA-Briefcase 1811, nearly tied with Opus 5.5 (1846 / 1822) and about 400 points above Sonnet 5

Benchmarks (Vendor-Reported)

Source: Anthropic’s announcement page anthropic.com/claude-sonnet-5-5 (September 28, 2026). Anthropic’s footnotes: the Opus 5.5 Terminal-Bench score is at xhigh effort and FrontierCode at max effort; some GPT-6 Sol scores are outdated because of a recent bug fix on that side. Retrieved September 30, 2026.

Independent Evaluation (Artificial Analysis)

On the same benchmark, the independent evaluation reports lower scores than Anthropic, but the ranking is the same:
A lower unit price does not always mean a cheaper task: Artificial Analysis measured Sonnet 5.5 at max effort using about 193,000 tokens per test task on average, the highest of any model it has tested. Sonnet 5.5 defaults to high effort. Running everything at xhigh / max inflates token usage and can cancel out the per-token savings. Start everyday tasks at medium or high rather than maxing out by default.

For Developers: Breaking Changes

Sonnet 5.5 keeps the Sonnet 5 request structure, but the following changes return errors or change the response shape. Check them before migrating:
top_p / top_k return errors: Anthropic’s documentation says Sonnet 5.5 rejects non-default temperature / top_p / top_k. In APIYI testing (2026-09-30), top_p and top_k are rejected on both the Anthropic-native and OpenAI-compatible endpoints with the message `top_p` is deprecated for this model. The status code shows as 429, but it is not rate limiting: retrying does not help, removing the parameter does. temperature returns 200 with any value but has no effect. Remove all three parameters in your client before switching.
Two more changes do not cause errors:
  • Progress text between tool calls now arrives in thinking blocks: with the default display: "omitted" their text is empty, so an interface that streams that text to users goes quiet between tool calls. Use between_tools or set thinking.display to get it back
  • Safety classifiers: a declined request returns HTTP 200 with stop_reason: "refusal", and stop_details names one of five categories (cyber / bio / frontier_llm / reasoning_extraction / general_harms). This is the first Sonnet with cyber and anti-distillation classifiers; higher-risk cyber requests fall back to Sonnet 5 on the provider side

Technical Specifications

APIYI Test Results

Tested on APIYI on September 30, 2026 (UTC+8): The last two rows show where cost comes from: the same model spends almost no thinking tokens on a simple question, and nearly ten thousand on a hard one at max. The effort level affects your bill far more than the unit price does.

Practical Applications

  1. Everyday coding: bug fixes, small repository changes, and terminal agent tasks; Terminal-Bench 4.0 shows the biggest gain
  2. Documents, slides, and spreadsheets: an area Anthropic highlights, with AA-Briefcase and GDPval-AA close to Opus 5.5
  3. High-frequency agents and support workflows: faster output and fewer tool-call steps suit latency-sensitive multi-turn tool use
  4. Replacing Sonnet 5 and some Opus 5.5 workloads: a same-price upgrade from Sonnet 5; well-scoped tasks on Opus 5.5 are worth evaluating on Sonnet 5.5 at about half the cost

Code Examples

Anthropic-Native Format

Turning Off Up-Front Thinking (Replaces disabled)

OpenAI-Compatible Format

Six-Step Migration Checklist from Sonnet 5

  1. Change the model name: claude-sonnet-5 → claude-sonnet-5-5
  2. Replace the thinking-off setting: change thinking: {"type": "disabled"} to between_tools, with effort no higher than high
  3. Remove forced tool use: set tool_choice to auto and use strict: true to keep tool input valid
  4. Drop sampling parameters: top_p / top_k return an error and temperature has no effect, so send none of them; stop using manual budget_tokens too and control thinking depth with effort
  5. Recalibrate effort: the same level does not produce the same amount of thinking as on Sonnet 5, so re-run your evaluation instead of carrying the old setting over
  6. Keep history append-only: pass thinking blocks back unchanged and do not edit system / tools / earlier messages between requests

Pricing and Availability

Pricing

Official prices (USD per million tokens): The Sonnet 5.5 column is also APIYI’s price: all four billing items match the official price, with no markup.
Model prices are aligned with the official website and may change with it; the table above is for reference only, and the Model Pricing tab in the top navigation is authoritative: Model Pricing.
Our pricing matches the official price, with no markup on the model. On top of that, some groups carry an additional discount, and recharge bonuses can be stacked. That is our own discount and separate from the model’s pricing.For actual charges, rely on the live data in the console’s cache billing details. The cache fields in the API’s usage response are not a billing reference.

Groups and Endpoints

The ClaudeCode group is for Anthropic-native clients such as Claude Code. Match the protocol to the group: use the ClaudeCode group for the Anthropic-native protocol and the default / svip groups for the OpenAI-compatible protocol. The ClaudeCode group carries an additional discount that stacks with recharge bonuses.
Sonnet 5.5 supports zero data retention and is not subject to the 30-day data retention requirement of the Fable series.

Stack with Recharge Promotions

APIYI recharge bonuses can lower your effective cost further. See Recharge Promotions.

Summary and Recommendations

Claude Sonnet 5.5 is a same-price upgrade: it costs exactly what Sonnet 5 did, beats Opus 5.5 on agentic coding, trails the flagship by only 2 points on knowledge work, and runs about 30% faster. For most everyday coding and document work, it is currently the best value in the Claude lineup. Recommendations:
  1. Sonnet 5 users: upgrade directly, but go through the six-step checklist first, especially disabled → between_tools, forced tool use, and sampling parameters
  2. Opus 5.5 users: evaluate well-scoped coding and document tasks on Sonnet 5.5 first at about half the cost; keep complex architecture decisions and long research chains on Opus 5.5
  3. Cost control: do not default to max effort; start at medium / high and raise it only if results call for it
Sources: Anthropic’s announcement page anthropic.com/claude-sonnet-5-5; Anthropic developer docs platform.claude.com/docs/en/models/sonnet-5-5/overview and whats-new-sonnet-5-5; Artificial Analysis independent evaluation (via Decrypt); reporting from SiliconANGLE, Unite.AI, Help Net Security, and others. The APIYI Test Results section comes from our own calls on September 30, 2026. APIYI pricing follows live platform data. Retrieved September 30, 2026.