Цены+7% бонус

What Is Claude Sonnet 5.5? Pricing, Coding Gains, and What Changed

Sonnet 5.5 keeps Sonnet 5’s official token rates. Explore coding results, Claude Code access, upgrade checks, and GPTProto’s 10% off API pricing.

What Is Claude Sonnet 5.5? Pricing, Coding Gains, and What Changed

Claude Sonnet 5.5 has the same API price per token as Sonnet 5. Anthropic nevertheless says the new model can cost less to finish a task. That sounds contradictory until you separate the price of each token from the number of tokens and tool calls a job actually takes.

The short answer: Claude Sonnet 5.5 is Anthropic's September 28, 2026 update to its Sonnet model for coding, document work, and other tasks with a clear goal. It accepts text and images, returns text, and is available in Claude Code and through the Claude API. Anthropic reports faster output and better results than Sonnet 5 on several evaluations. Those results are reasons to test an upgrade, not guarantees for every codebase or workflow. The Claude Sonnet 5.5 model page on GPTProto is the place to check its platform availability and try it when the listing is active. Anthropic's announcement sets out the release and its claims.

Содержание

What is Claude Sonnet 5.5?

Sonnet is the middle tier of Claude's current lineup. Sonnet 5.5 is meant to handle defined, repeatable work quickly: implementing a specified feature, fixing a bug, reviewing a document against source material, or producing a presentation from a brief. Anthropic positions Opus 5.5 for harder, open-ended work that needs sustained judgment. The distinction matters more than the version number. A high score on one coding test does not make Sonnet the better choice for every architecture decision.

Specification Claude Sonnet 5.5
Release date September 28, 2026
Claude API model ID claude-sonnet-5-5
Context window 1 million tokens
Maximum output 128,000 tokens on standard requests
Modalities Text and image input; text output
Thinking Adaptive, enabled by default

These are Anthropic's published specifications. A 1-million-token context window is substantial, but it is not a new advantage over Sonnet 5, which also supports 1 million tokens. Nor does a large window prove that every detail in a long input will be used correctly. Sonnet 5.5 also keeps Sonnet 5's tokenizer, so an identical piece of text has the same token count on those two models. Anthropic's migration documentation makes that last point explicit.

What changed from Sonnet 5?

The useful change is how much work Sonnet 5.5 can complete for a given amount of time and computation. Anthropic says it generates output more than 30% faster than Sonnet 5. In its Terminal-Bench 4.0 agentic coding evaluation, Sonnet 5.5 scored 70.6%, compared with 10.3% for Sonnet 5. It also reports that the newer model batches tool calls more often, reducing the steps needed for some coding tasks. These are Anthropic's results, measured in particular test setups.

There is an outside check, though it asks a different question. On CursorBench 4.0, Sonnet 5.5 at High effort scored 47.8% at a reported $1.67 per evaluated task; Sonnet 5 at High scored 30.8% at $3.48. That supports an improvement in Cursor's coding environment. It does not predict the exact saving for an API customer using a different agent, prompt, or repository.

The upgrade is therefore more interesting than a routine version bump for coding teams. The tradeoff is that a strong benchmark result still needs to survive your own tests. If you want the earlier model's starting point, see our Sonnet 5 guide and Sonnet 5 API page.

Sonnet 5.5 pricing: same list rate, potentially fewer tokens per task

At Anthropic's published API rates, Sonnet 5.5 costs the same per million tokens as Sonnet 5. GPT Proto's Sonnet 5.5 rate is 10% below those input and output rates:

Per 1 million tokens Anthropic list rate GPT Proto Sonnet 5.5
Input $2.00 $1.80
Output $10.00 $9.00

Anthropic lists cache reads at $0.20 per million tokens; cache writes and other charges have their own rates. Check the GPT Proto model page for its live billing details before estimating a cached workload. The Anthropic rates and the fact that they match Sonnet 5 come from the release announcement. The GPT Proto input and output rates above reflect its 10% model discount.

Here is a deliberately simple calculation. If one request uses 100,000 ordinary input tokens and 20,000 output tokens, with no cache or tool charges, the token cost would be $0.40 at Anthropic's list rates and $0.36 at GPT Proto's stated rates. That is a comparison for identical token usage. Separately, Anthropic says Sonnet 5.5 used fewer tokens to complete typical tasks in its testing, reducing cost per finished task by up to 30% versus Sonnet 5. Your result depends on the task and effort setting; 30% is not a cut to the token tariff, and it should not be added mechanically to GPT Proto's 10% rate discount. See GPT Proto pricing for account-level billing information.

Is Sonnet 5.5 good for coding and available in Claude Code?

Yes. Anthropic's Claude Code model guide lists Sonnet 5.5. Use /model inside Claude Code to select it, or start a session with claude --model claude-sonnet-5-5. Claude Code and the Claude apps default to Medium effort for this model; the Claude API defaults to High. Effort controls how long the model works through a problem, affecting quality, time, and billed output. The defaults therefore matter when someone quotes a score obtained at Max effort as though it describes an ordinary Claude Code session.

In one developer's early head-to-head test, the author ran ten personal coding tasks three times per configuration at High effort in Claude Code and pi-agent. Sonnet 5.5 used fewer turns and output tokens than Sonnet 5 in that setup. The author also noted that newer models were already passing nearly all the tasks, limiting what the test could say about difficult work. It is a useful observation about workflow, not a universal win rate.

For well-specified implementation and bug fixing, I would test Sonnet 5.5 first. For an ambiguous architecture change or a review where a subtle miss is expensive, also test Claude Opus 5.5. Anthropic itself says Opus remains stronger at complex, open-ended work, despite Sonnet 5.5's high scores on individual benchmarks. Its release discussion draws that boundary clearly.

Should you upgrade an existing Sonnet 5 integration?

Probably, if your own evaluation improves. But changing the model string is only the first step for an existing API integration. Anthropic's Sonnet 5.5 migration guide documents request settings that can return a 400 error after the switch:

If your Sonnet 5 integration does this Check before switching
Sends thinking: {"type":"disabled"} Sonnet 5.5 uses between_tools for its lowest thinking setting.
Forces a particular tool with tool_choice Sonnet 5.5 rejects forced tool and any choices; review the documented auto approach.
Uses an older computer-use tool version The accepted toolset differs by platform; follow the migration table for yours.
Assumes the first response block is text Read blocks by type and preserve thinking blocks in tool loops.

There is also a quieter change: effort levels have been recalibrated. High on Sonnet 5.5 need not consume the same amount of work as High on Sonnet 5. Anthropic recommends rerunning your own evaluations for quality, latency, and cost rather than carrying over a previous setting. This is particularly relevant if your agent makes repeated tool calls or if you parse model output in a fixed format.

For a basic GPT Proto chat request, use the model page's Try this model action and API Usage panel to confirm that Sonnet 5.5 is active and that the displayed model ID is claude-sonnet-5-5. Once it is active, the standard GPT Proto chat-completions request has this shape:

curl --request POST 'https://gptproto.com/v1/chat/completions' \
  --header "Authorization: Bearer $GPTPROTO_API_KEY" \
  --header 'Content-Type: application/json' \
  --data '{
    "model": "claude-sonnet-5-5",
    "messages": [{"role": "user", "content": "Summarize the risks in this project brief in three sentences."}]
  }'

Set GPTPROTO_API_KEY to your own key before running it. GPT Proto shows this request format on its existing Claude model pages; the Sonnet 5.5 model page is the final source for its live model string and availability. Provider-specific Claude Messages API options need separate compatibility checks when used through an OpenAI-compatible endpoint.

How does Sonnet 5.5 compare with other models?

An explainer cannot settle every “versus” search with a single leaderboard. The comparison changes with the task, agent setup, effort, and price being measured. These are three useful but narrow reference points:

Question What a comparable test says What it cannot establish
Sonnet 5.5 vs GPT-6 Astra On Vals AI's Code Migration evaluation, Sonnet 5.5 scored 69.83% and Astra 67.74%. Which model is better for all coding, research, or writing tasks.
Sonnet 5.5 vs GLM-5.3 On CursorBench 4.0, Sonnet 5.5 High scored 47.8% and GLM-5.3 High 38.0%. Performance or cost in your own agent and repository.
Sonnet 5.5 vs Claude Fable 5.1 On the Vals Index, Sonnet 5.5 scored 69.22% and Fable 5.1 68.83%. A reliable overall winner from a small composite-score gap.

The Astra and Fable results come from Vals AI's evaluation, which says it used Max effort for most tests. The GLM result comes from Cursor's own benchmark. Do not combine the scores from these two sources into one ranking. If a specific alternative is cheaper or already fits your stack, the relevant test is whether Sonnet 5.5 completes your task more accurately or with less total spend.

Who should try Sonnet 5.5?

Start with work that has a visible finish line: a defined code change, a bug with a failing test, a document with a source packet, or a repeatable analysis. Run the same small batch of real tasks through Sonnet 5 and 5.5 at explicit effort settings. Track whether the answer is correct, whether the model checked its work, how long it took, and how many tokens were billed. That gives you a much better upgrade decision than selecting the largest launch-day percentage.

Sonnet 5.5 is a strong candidate for those bounded jobs. It is less safe to assume that one benchmark makes it a substitute for Opus on open-ended work, or that Anthropic's maximum reported saving will appear on every bill. When the Sonnet 5.5 API page on GPT Proto is active, you can test it alongside the other models linked above with the same account and compare your own results.

FAQ

When was Claude Sonnet 5.5 released?

Anthropic released it on September 28, 2026. Its Claude API model ID is claude-sonnet-5-5

Is Sonnet 5.5 available in Claude Code?

Yes. Select it with /model or start Claude Code with claude --model claude-sonnet-5-5

Is Sonnet 5.5 cheaper than Sonnet 5?

Anthropic charges the same $2 input and $10 output per million tokens for both. It reports lower cost per completed task in some tests because Sonnet 5.5 uses fewer tokens. On GPTProto, the stated Sonnet 5.5 input and output rates are $1.80 and $9.00 per million tokens, 10% below Anthropic's list rates.

Is Sonnet 5.5 better than GPT-6 Astra?

It scored higher on Vals AI's Code Migration test, but that one result does not answer the broader question. Test both on the coding or agent workflow you intend to run, at documented settings and with the full task cost recorded.

Похожие статьи

Ещё блоги
What Is Kling 4.0? Preview Leaks, Release Status, and What You Can Use Now

What Is Kling 4.0? Preview Leaks, Release Status, and What You Can Use Now

Status checked: September 28, 2026. This article will need an update when Kling publishes a 4.0 announcement or API documentation. The first question about Kling 4.0 is surprisingly practical: can you make a video with it today? I cannot point to an official public Kling 4.0 release or a documented 4.0 API model ID yet. What people are calling “Kling 4.0” is a possible next generation of Kuaishou's Kling video models, discussed after screenshots and a feature list circulated online. Kling's public site still presents the 3.0 series as its current lineup. That distinction matters if you are planning a product video or wiring an API into a production workflow. A setting seen in a preview screen is a clue about development. It is not a promise about the model you can buy, call, or compare today.

Schuyler Stacy | 2026-09-28

6 Best LLM API Providers in 2026: Multi-Model Platforms Compared

6 Best LLM API Providers in 2026: Multi-Model Platforms Compared

Choosing an LLM API provider is no longer the same as choosing a model. The same open-weight model can be available from several platforms, yet the real service you receive may differ in latency, throughput, context limits, tool calling, caching, error behavior, and price. The lowest listed token price can cost more in production if cache hits are unreliable or retries are frequent. An “OpenAI-compatible” endpoint may also accept basic chat requests while rejecting fields your application needs. We compared six multi-model LLM API providers across aggregators, managed cloud platforms, and inference specialists. First-party APIs such as OpenAI and Anthropic remain useful baselines, but they do not offer the same cross-vendor access. One Key for Your Team

Tiffany Layne | 2026-09-21

Claude Opus 5.5 vs Sonnet 5.5: Which Is Worth the Cost for Coding and Agents?

Claude Opus 5.5 vs Sonnet 5.5: Which Is Worth the Cost for Coding and Agents?

Claude Sonnet 5.5 costs half as much as Claude Opus 5.5 per million input and output tokens. On GPTProto, both models are listed at 10% below Anthropic's standard input and output rates. That makes the first decision easy for routine, well-defined work: start with Sonnet. For an ambiguous codebase change or a review where a missed defect is expensive, test Opus before deciding that its higher rate is wasteful. The awkward part is that the lower token rate does not guarantee a lower bill for every completed task. I would choose between them by looking at accepted results, time, and total usage on the same work. A model that finishes cheaply but needs a second pass can lose its apparent price advantage. At the highest effort setting in one independent evaluation, Sonnet actually spent more per task than Opus. That is a useful warning, though it does not describe every developer workload. Get Sonnet 5.5 API

Schuyler Stacy | 2026-09-29