Research index
Model releases/

Claude Sonnet 5.5 API: $2/$10 Pricing and Setup

Claude Sonnet 5.5 is $2/$10 per 1M tokens. API id claude-sonnet-5-5. Not on RouterPlex yet. Register for 60+ live models and use it once listed.

Written byRouterPlex
Reading time6 min
Claude Sonnet 5.5 API: $2/$10 Pricing and Setup
Claude Sonnet 5.5 API: $2/$10 Pricing and Setup

Claude Sonnet 5.5 is Anthropic's second Claude 5.5 model, released 28 September 2026. The Claude API id is claude-sonnet-5-5. Claude Sonnet 5.5 API pricing is $2 per 1M input tokens and $10 per 1M output tokens, the same sticker as Sonnet 5, with cache reads at $0.20 per 1M.

Sonnet 5.5 is not on RouterPlex yet. Create a RouterPlex account to use the current 60+ catalog models today, including Claude Sonnet 5 at those same $2/$10 rates. The new id works here once the marketplace lists it. One prepaid key, 0% markup on vendor list price, a $0 top-up fee, and a hard spend limit that stops at $0.

Anthropic launch card for Claude Sonnet 5.5, 28 September 2026.
Anthropic launch card for Claude Sonnet 5.5, 28 September 2026.

Sources: Anthropic's Claude Sonnet 5.5 announcement (28 September 2026), the Sonnet 5.5 overview, the migration guide, and the system card, checked 28 September 2026. Benchmark figures below are Anthropic's reported results. Rates change. Confirm the vendor page before you lock a budget.

Claude Sonnet 5.5 API pricing #

USD per 1M tokens on Anthropic's platform. The announcement table's cache-write cell of $2.50 is the 5-minute rate. The overview also publishes a 1-hour cache write.

LinePer 1M tokens
Input$2.00
Output$10.00
Cache read$0.20
5-minute cache write$2.50
1-hour cache write$4.00

Anthropic's Batch API is 50% off input and output on that platform. Prompt caching up to 90% is also an Anthropic platform claim. Neither discount is a RouterPlex bill, because this model is not listed here.

A 40,000-token input and 2,000-token output, with no cache, is $0.08 + $0.02 = $0.10 on this sheet. That is the same arithmetic as live claude-sonnet-5 on RouterPlex. The 5.5 id is a different model.

The minimum cacheable prompt is 512 tokens.

What Sonnet 5.5 is #

Claude Sonnet 5.5
Claude API idclaude-sonnet-5-5
Bedrock idanthropic.claude-sonnet-5-5
Vertex, Foundry, Claude Platform on AWSclaude-sonnet-5-5
Released28 September 2026
Context / max output1M / 128K (Batch beta 300K with output-300k-2026-03-24)
LatencyFast
ThinkingAdaptive. Lowest setting between_tools turns off up-front thinking
Default effortHigh on the Claude Platform. Medium in the apps and Claude Code
Knowledge cutoffJune 2026
RetirementNot sooner than 28 September 2027
StatusActive
On RouterPlexWhen listed

Anthropic positions it for well-scoped everyday work next to Opus 5.5, which stays the stronger model for open-ended jobs. Zero data retention is available, same as Opus 5.5 and Sonnet 5, on Anthropic's terms.

If you currently disable thinking, switch to between_tools before moving. It works at high effort or below. Non-default temperature, top_p, or top_k returns HTTP 400.

Vendor benchmarks, 28 September 2026 #

These scores are Anthropic's. FrontierCode 1.1 Main is reported twice for Sonnet 5.5: 46.2% at Max and 52.1% at Xhigh. The other columns on the Xhigh row are blank on the announcement. Publish both effort labels. Do not treat 46.2% as the only FrontierCode score.

BenchmarkSonnet 5.5Sonnet 5Opus 5.5GPT-6 Sol
Terminal-Bench 4.070.6%10.3%66.4%not reported
FrontierCode 1.1 Main, Max46.2%42.4%54.4%49.3%
FrontierCode 1.1 Main, Xhigh52.1%
CursorBench 4.055.534.157.8
GDPval-AA v2.11844144918461487
AA-Briefcase v1.11811135918221483
HLE, with tools64.5%54.9%67.7%
OSWorld 2.1, partial80.157.081.8
Chartography, no tools61.615.664.453.6

Terminal-Bench and OpenAI did not publish a GPT-6 Sol number, so Anthropic's chart uses GPT-5.6 Sol there. FrontierCode penalizes out-of-scope changes. Anthropic's footnote says that at Max, Sonnet 5.5 more often ran Claude Code's code-review skill, and in two cases Cognition examined that produced a timeout or extra edits, which lowered the score. Artificial Analysis ran GDPval-AA and AA-Briefcase on a pre-release deployment with a structured-output bug. Anthropic says to expect a small effect.

Other vendor claims from the same page, not RouterPlex measurements: at High effort on FrontierCode, about 10 points above Sonnet 5 at the same setting and about one fifteenth of the cost per task. At Medium effort, the apps default, Terminal-Bench far exceeds Sonnet 5's best for less than a tenth of the cost per task. Output is more than 30% faster, and cost per task is up to 30% lower because the model uses fewer tokens.

Still from Anthropic's 28 September 2026 Sonnet 5.5 page for the HTML murmuration prompt. The frame reads Output 4,158 tokens, Flying 12.5 s. The still does not name the model.
Still from Anthropic's 28 September 2026 Sonnet 5.5 page for the HTML murmuration prompt. The frame reads Output 4,158 tokens, Flying 12.5 s. The still does not name the model.

Daniel Vogel, COO at Epic Games, said early testing cleared the quality bar expected from a higher-tier model on a system design audit and a data flow review, including tens of thousands of lines and multi-hour tasks. Curtis Allen, Principal Engineer at Slack, said Slackbot offline evals improved with about 14% fewer output tokens and fewer steps, without prompt changes. Anthropic also says this is the first Sonnet to beat Pokemon Red from screenshots only.

Safeguards #

This is the first Sonnet with cyber safeguards and reasoning-extraction classifiers. Higher-risk cyber tasks fall back to Sonnet 5. Biology safeguards match Sonnet 5. Routine software work is the intended path. Details are in the system card.

Call Claude Sonnet 5 today #

claude-sonnet-5 is the live RouterPlex route at $2 / $10 per 1M. It is the previous Sonnet, not 5.5.

bash
curl https://api.routerplex.com/v1/chat/completions \
-H "Authorization: Bearer $ROUTERPLEX_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "claude-sonnet-5",
"messages": [
{"role": "user", "content": "Name the riskiest assumption in this migration plan."}
]
}'

The same key reaches the rest of the catalog through the OpenAI-compatible API. Compare gateways on the OpenRouter comparison. Opus 5 is live at $5 / $25. Fable 5 is live at $10 / $50.

On RouterPlex, once Sonnet 5.5 is listed #

When the row exists, the contract is vendor list price, 0% markup, prepaid balance, and a hard per-key budget. The request stops at $0. This page will name the RouterPlex route that day. Until then, paste claude-sonnet-5, not claude-sonnet-5-5.

Create a RouterPlex account if you want the key ready.

Common questions

Frequently asked questions

How much does the Claude Sonnet 5.5 API cost?

Anthropic lists Claude Sonnet 5.5 at $2 per 1M input tokens and $10 per 1M output tokens. Cache reads are $0.20 per 1M. A 5-minute cache write is $2.50 per 1M and a 1-hour cache write is $4 per 1M. Those are Anthropic platform rates, checked 28 September 2026. RouterPlex does not bill Sonnet 5.5 yet.

What is the Claude Sonnet 5.5 model ID?

On the Claude API the ID is claude-sonnet-5-5. Amazon Bedrock uses anthropic.claude-sonnet-5-5. Vertex AI, Microsoft Foundry, and Claude Platform on AWS use claude-sonnet-5-5. There is no RouterPlex model ID until the catalog lists it.

Is Claude Sonnet 5.5 on RouterPlex?

Not yet. Register at routerplex.com/sign-up to use the current 60+ catalog models, including claude-sonnet-5 at the same $2/$10 token rates. Sonnet 5.5 can be called here once it is live on the marketplace.

What is the Claude Sonnet 5.5 context window?

Anthropic documents a 1M-token context window and a 128K maximum output. The Message Batches API beta allows up to 300K output tokens with the output-300k-2026-03-24 header. Knowledge cutoff is June 2026. Retirement is not sooner than 28 September 2027.

How does Sonnet 5.5 compare with Sonnet 5 on price?

The token rates match Sonnet 5: $2 input and $10 output per 1M. Anthropic says Sonnet 5.5 typically uses fewer tokens, up to 30% less cost per task in their tests, and generates output more than 30% faster. That cost-per-task claim is Anthropic's, not a RouterPlex measurement.

Can I turn thinking off on Sonnet 5.5?

Adaptive thinking is on by default. The lowest setting is between_tools, which turns off up-front thinking and works at high effort or below. Setting temperature, top_p, or top_k to a non-default value returns a 400 error.

Run the smallest paid test.

Add $5, cap the key, and verify the result with your own workload. No subscription, and credit never expires — a first top-up of $25+ is matched with $25 extra.