요금제+7% 보너스

A Lower-Cost API forCLI Coding Agents

Keep the coding CLI you already use. Compare GPTProto rates for supported models, connect through the protocol your client expects, and test the setup on one task before moving a full workflow.

제공 모델
Set up GPTProto for Claude Code.

1. In ~/.claude/settings.json, under "env" set:
     ANTHROPIC_BASE_URL = "https://gptproto.com"
     ANTHROPIC_API_KEY  = "sk-gptproto-<your key>"
   Keep it exactly like that - no /v1, no trailing slash. Claude Code appends /v1/messages itself, so a /v1 here becom...
Get API Key
// 바꾸는 이유

Long coding sessions makeevery API detail count.

A refactor can involve many model calls. Before changing providers, check the rate for your model, the client protocol, and how failed requests are handled.

01

The bill grows across long sessions.

Repo-wide edits and repeated test loops can use more input, output and cached tokens than a short chat. Compare the rate for the model and token types you actually use.

Compare model rates →
02

A 429 can interrupt a run.

Rate limits and upstream errors can break a coding session. Check your client's retry behavior and test GPTProto's supported routes before moving important work.

How requests are handled →
03

A mismatched endpoint wastes time.

Claude Code, Codex CLI and OpenCode do not all use the same API shape. Choose your client first, then follow its matching endpoint and model setup.

Top up $100+ → wan-3.0, free & unlimitedSep 29 – Oct 10 · Limited quota · First come, first served
Top up now
// Live rates

Compare the coding modelsyou already use.

Check GPTProto's current input, output and cache-read rates beside the provider's published API rates. Discounts vary by model; availability in a specific CLI depends on its protocol and supported features.

The models Claude Code, Codex CLI and their peers ship with — the ones you are most likely already sending. Rates are per 1M tokens, discounted from each vendor's list.

도구 호출긴 컨텍스트Cache-aware pricing
API 키 받기
모델가격: GPTProto공식 대비OpenRouter 대비컨텍스트모달리티안정성작업
Claude Opus 5Claude
$4.50 / $22.50$0.45 캐시 읽기 · 1M 토큰당−10%−15%1M→사용해 보기
Claude Sonnet 5Claude
$1.80 / $9.00$0.18 캐시 읽기 · 1M 토큰당−10%−15%1M→사용해 보기
Claude Haiku 4.5 (Build 20251001)Claude
$0.90 / $4.50$0.09 캐시 읽기 · 1M 토큰당−10%−15%200K→사용해 보기
GPT 5.2 CodexOpenAI
$1.23 / $9.80$0.12 캐시 읽기 · 1M 토큰당−30%−34%400K→사용해 보기
GPT 5.1 Codex MaxOpenAI
$0.88 / $7.00$0.09 캐시 읽기 · 1M 토큰당−30%−34%400K→사용해 보기
GPT 5.4 MiniOpenAI
$0.60 / $3.60$0.06 캐시 읽기 · 1M 토큰당−20%−24%400K→사용해 보기
GPT 5 NanoOpenAI
$0.035 / $0.28$0.00 캐시 읽기 · 1M 토큰당−30%−34%400K→사용해 보기
// START WITH YOUR WORKFLOW

Find a starting balance for
your coding workload.

Choose the kind of CLI work you do and how often you run it. This is a starting point; your actual spend depends on model choice and token usage.

01 / What is your CLI doing?
02 / How much do you burn a month?

Batch work and long-context runs burn credits faster — the recommendation adjusts automatically as you switch workload or usage level.

$100
Coding agent · 보통 — 매주 — Repo-wide edits and test loops — long tool-call chains where input and cache tokens dominate the bill.
$104.00 수령 · +4% 보너스 크레딧
16.5M–33M 토큰
공식 대비 절약$30.00
OpenRouter 대비 절약$37.15
Top up $100+ → wan-3.0, free & unlimitedSep 29 – Oct 10 · Limited quota · First come, first served
Top up now
// The result

Keep your CLI.Use the setup it actually needs.

GPTProto exposes different API routes for different clients. Select a supported model, configure the matching endpoint and key, then confirm a short coding task works as expected.

Before · vendor endpoint
✕Claude Code → api.anthropic.com
✕Codex CLI → api.openai.com/v1
✕Two keys, two balances, two invoices
✕Each vendor’s own rate limit
After · gptproto
✓Claude Code → gptproto.com (no /v1)
✓Codex CLI → gptproto.com/v1
✓One key, one balance, one bill
✓10–30% below either vendor’s list price

2 lines changed · 0 code rewrites · 200+ models on one key · model names stay the same.

// The mechanism

Keep your CLI.
Use the setup it actually needs.

GPTProto exposes different API routes for different clients. Select a supported model, configure the matching endpoint and key, then confirm a short coding task works as expected.

01

Two API shapes, clearly labeled.

See which client uses the Anthropic-compatible Messages route and which needs the OpenAI Responses route. Copy only the setup for your CLI.

02

One account for the models you use.

Manage supported model access and API spending in one GPTProto account. Check the rate and available features before changing a model in your CLI.

Explore supported models →
03

Handle interruptions explicitly.

Where a supported model has alternative upstream routes, GPTProto may reroute eligible requests. A request can still fail; keep retry and backoff in your client.

// AI API 절약 계산

Estimate your codingagent API bill.

Select a model and enter your monthly input, output and cached-input tokens. We compare the same workload against that model's published provider API rates.

텍스트 · 100만 토큰당
토큰 / 월
$816
Claude 대비 연간 절약 및 OpenRouter(추정)
년월일
Claude$6,000$500.00$16.44
OpenRouter (est.)$6,330$527.50$17.34
GPTProto$5,400$450.00$14.79
Top up $100+ → wan-3.0, free & unlimitedSep 29 – Oct 10 · Limited quota · First come, first served
Top up now
// 사용자 이야기

우리 말만 믿지 마세요.사용자 이야기를 보세요.

실재하는 공개 계정의 글입니다. 여기 있는 것은 만들어낸 것이 아닙니다.

BL
Blogstra
Reddit · r/Bloggers
#Routing

For a team already juggling DeepSeek, Kimi, Qwen, or other providers, a shared routing layer can be reasonable if it removes duplicated work and the fallback behavior is tested. GPTProto is one implementation of this approach.

K
K
X (Twitter) · @ChillaiKalan__
#Cost

I've been testing GPTProto recently, and it's honestly made my creative workflow much simpler. Instead of paying for multiple subscriptions, I can access several leading AI models from one place.

SN
Sanskriti Naruka
X (Twitter) · @snskritinaruka
#Video

I turned this single prompt into a cinematic fantasy video using GPTProto. "Continuous 15-second cinematic shot, 4K resolution, hyper-realistic dark fantasy photorealism…"

CF
Caden Flux
X (Twitter) · @Caden_Flux
#Video

I challenged myself to create a cinematic AI cooking short in just 15 seconds. I used GPTProto to bring the entire workflow together, from image generation to video, all in one place.

Z
Zara
X (Twitter) · @ZaraIrahh
#Creative

Made with Seedance 2.0 + GPT Image 2 on GPTProto. A Pixar-style commercial with the perfect glow.

MO
Many-Operation2625
Reddit · r/LLMDevs
#Integration

What worked for me was pointing the cloud connections at GPTProto so the frontend only sees one endpoint and I just change the model name to swap. I am not rebuilding a connection from scratch every time.

// 요금

Pay for the model calls
your CLI makes.

Start with a small balance, test a real session and scale up when you know your usage. Model prices and any top-up bonuses are shown separately.

$10
정액 충전 · 보너스 없음

Enough to point one CLI at the endpoint and run a real session. Top up $20+ anytime to unlock bonus credits.

최대 절약 · +6%
$1,000
+6% 보너스 크레딧

Built for a fleet of CLIs and agent runners. Bonus credits stack on top of already-discounted model pricing, so the whole balance prices 10–30% under list.

GPT 5.2 Codex에서 월 1억 토큰을 쓰면, GPTProto 요금으로는 공식보다 연간 약 $630 덜 듭니다 — 보너스 크레딧 적용 전.
// 자주 묻는 질문

Before you changeyour CLI setup.

The answers that matter when you compare model rates and test a new endpoint.

01Can I use GPTProto with Claude Code, Codex CLI and OpenCode?+

These clients have different provider settings and API requirements. GPTProto offers Anthropic-compatible Messages and OpenAI Responses routes for supported models. Choose the guide for your client and run a short test; a model listed in the catalog is not automatically compatible with every CLI feature.

02Why are the Claude Code and Codex CLI setup steps different?+

Claude Code uses an Anthropic-compatible Messages route when configured with a compatible gateway. Codex CLI's custom providers use the Responses API. Each client needs its own base URL, authentication settings and supported model ID; follow the matching guide instead of copying one CLI's config into another.

03Can I keep my prompts, tools and repo setup?+

Start by changing the provider configuration, then test the tasks you rely on. Prompt and repository files may remain the same, but streaming, tool calls, model options and output formats can behave differently. Check a real edit-and-test loop before switching your daily workflow.

04Can I use the same model ID I use with the official provider?+

Use the model ID shown in GPTProto's catalog for the route you selected. Model availability and supported features may differ by endpoint. Check the model page and your CLI setup guide before starting a long run.

05Does switching endpoints remove rate limits or prevent 429 errors?+

No API can promise that. Limits and upstream failures can still occur. GPTProto may route eligible requests through alternative channels where supported, but your CLI or application should still handle retries and backoff. Test behavior with your chosen model before moving important work.

06Will my CLI subscription become cheaper?+

This page compares model API usage, billed according to tokens and applicable model rates. It does not change a subscription or bundled usage you may have with the CLI vendor. Check whether your current workflow uses a subscription, a vendor API key or both before comparing costs.

07How are savings calculated?+

We apply the selected model's input, output and, where relevant, cached-input rates to the monthly usage you enter. The comparison uses the same token quantities at the provider's published API rates. Any top-up bonus appears as a separate line so you can see the model discount on its own.

08Do bonus credits apply to already discounted model rates?+

If a top-up tier includes bonus credits, the checkout page shows the total credit you receive. Usage is charged at the displayed GPTProto model rate. Bonus amounts, eligibility and validity depend on the current offer; check them before paying.

09Can I switch back to my previous provider?+

Yes, you can restore your previous provider settings. Keep a copy of the original configuration and test both paths with a small task. Unused GPTProto credit remains subject to the account terms shown at checkout.

10Does GPTProto itself have usage limits?+

Your key, account, chosen model and upstream route can each have limits. Review the limits shown for your account and model, and contact support before a workload that needs guaranteed capacity. The page should not promise unlimited calls without a documented commitment.

// 시작

Your next refactorshouldn't run at list price.

브라우저에서 이미지와 영상을 만들거나, API 키 하나로 텍스트, 이미지, 영상 모델을 앱에 넣을 수 있습니다. 할인 요금도 확인할 수 있습니다.

  • ✓Both protocols, one host
  • ✓200+ models on one key
  • ✓10–30% below official
  • ✓높은 가동률, 자동 페일오버
✓클립보드에 복사했습니다