1M-Token Context
Read large repositories, long documents, or extended agent histories in one request. Keep retrieval and validation in your workflow: a larger window does not guarantee that every detail will be used correctly.
Chat, coding agents & document work. Priced per 1M tokens — input, cached input and output are billed separately. GPTProto is 10% below official rates.
| シナリオ | Claude リスト | OpenRouter | GPTProto | 月間の節約 |
|---|---|---|---|---|
| Personal10M トークン / 月(キャッシュ 4.8M) | $27.36 | $28.86 | $24.62 | −$2.74≈ $32.83 / 年 |
| Team100M トークン / 月(キャッシュ 48M) | $273.60 | $288.65 | $246.24 | −$27.36≈ $328.32 / 年 |
| Business500M トークン / 月(キャッシュ 240M) | $1368.00 | $1443.24 | $1231.20 | −$136.80≈ $1641.60 / 年 |
Use Sonnet 5.5 for well-scoped development and knowledge tasks. It pairs 1M-token context with up to 128K output; tune effort and measure cost per completed task.
1M-Token Context
Read large repositories, long documents, or extended agent histories in one request. Keep retrieval and validation in your workflow: a larger window does not guarantee that every detail will be used correctly.
Faster Coding Iterations
Anthropic reports output generation more than 30% faster than Sonnet 5. Test the improvement on your own prompts, tool calls, and acceptance checks before moving a production workload.
Adaptive Effort for Agents
Adaptive thinking is enabled by default. Use lower effort for routine, clearly defined steps and higher effort for difficult decisions; output and thinking tokens both affect the final bill.
One Key Across Models
Use your GPTProto account to compare Sonnet 5.5 with other available text models. Route routine work and harder reviews separately, and check each model's current price and supported features before switching.
Claude Sonnet 5.5 is Anthropic's September 28, 2026 update to Sonnet 5. It accepts text and images and returns text. Anthropic positions it for coding, analysis, visual understanding, and agentic tool use; Opus 5.5 remains better suited to complex, open-ended work.
It keeps Sonnet 5's $2 input / $10 output per million tokens on Anthropic's direct API. Anthropic reports output generation more than 30% faster and up to 30% lower cost per task through fewer tokens, rather than a lower per-token rate. GPTProto's actual charge appears in the fixed pricing module.
| Specification | Claude Sonnet 5.5 |
|---|---|
| Developer / release | Anthropic / September 28, 2026 |
| Anthropic API model ID | claude-sonnet-5-5 |
| Input → output | Text and images → text |
| Context window / maximum output | 1M tokens / 128K tokens |
| Thinking / default Claude API effort | Adaptive / high |
| Official direct API list rate | $2 input / $10 output per 1M tokens |
| GPTProto rate | See the live pricing module above |
Repository changes and code review. Give an agent a defined issue, relevant files, tests, and acceptance criteria. Let the application supply tools and validate the result. Escalate difficult architecture decisions.
Long-document analysis. Supply specifications, support records, or reports and request structured answers with source references. The 1M-token context window adds capacity, but extraction checks still matter.
Image-aware workflows. Send screenshots or charts with text to explain a UI defect or extract information. This is image understanding with text output, not image generation. Verify GPTProto's image-input route before deployment.
These are Anthropic direct API prices and Anthropic-reported evaluation results. They are provided for model selection; the GPTProto pricing widget is the source for GPTProto billing.
| Decision factor | Sonnet 5 | Sonnet 5.5 | Opus 5.5 |
|---|---|---|---|
| Official input / output per 1M tokens | $2 / $10 | $2 / $10 | $4 / $20 |
| Context / maximum output | 1M / 128K | 1M / 128K | 1M / 128K |
| CursorBench 4.0 (Anthropic report) | 34.1% | 55.5% | 57.8% |
| Starting point | Existing integrations that still need evaluation | Well-scoped coding, analysis, and frequent agent steps | Complex, open-ended work and sustained judgment |
The same per-token rate does not mean the same bill. Log tokens, cache usage, tool calls, retries, and accepted results. Benchmarks help orient model selection, but effort and task setup change total cost. Route repeatable steps to Sonnet 5.5 and evaluate Opus 5.5 for harder decisions.
Replacing the model string alone can break an integration. Check the ID and request format in GPTProto's API Usage tab. If the route exposes Anthropic's native controls, test these changes before moving live traffic:
Replace claude-sonnet-5 with the supported Sonnet 5.5 ID. If you used thinking: {"type": "disabled"}, use thinking: {"type": "between_tools"} at low, medium, or high effort; disabled returns a 400 error on the new model.
Remove forced tool_choice values any and tool. Use auto, validate tool input, and test whether the model calls a required tool in your workflow.
Parse response blocks by type and preserve thinking blocks when continuing a conversation. Review streaming UIs that previously displayed text between tool calls.
Recheck max_tokens, effort, and non-default sampling parameters. Thinking consumes output tokens, while custom temperature, top_p, or top_k values can return a 400 error.
Test the exact payload on your chosen GPTProto route; native Anthropic controls may differ across routes.
Start with Sonnet 5.5 for repeatable code fixes, document work, and agent steps with clear success criteria. Compare Claude Opus 5.5 for extended judgment. When migrating from Claude Sonnet 5, measure cost per accepted result. Explore the model catalog to test other providers with one account.