API access is not available yet. This page tracks release status.
by Kling·API access planned
Every model's modalities, price and discount in one list — call any of them on the key you already use.
API access is not available yet. This page tracks release status.
by Kling·API access planned
Access GPT 6.1 Sol through GPTProto for coding, document analysis, and multi-step agent work. OpenAI positions the model near Astra for demanding tasks at lower standard token rates. Use one GPTProto key and balance to compare 200+ models across providers.
by OpenAI·$1.60 in / $8.00 out per 1M tokens
Access Anthropic's Claude Sonnet 5.5 through GPTProto for coding, agent workflows, document analysis, and visual understanding. Compare its live token rate with other models using one API key and shared balance.
by Claude·$1.80 in / $9.00 out per 1M tokens
Access GPT-6 Luna on GPTProto for high-volume chat, classification and structured extraction. Run focused tasks with adjustable reasoning and a 1.05M-token context window. Check the live rate and test the model using your GPTProto API key.
by OpenAI·$0.080 in / $0.40 out per 1M tokens
Access GPT-6 Sol on GPTProto for demanding coding, code review and multi-step agent work. Use its 1.05M-token context and adjustable reasoning to handle substantial task material. Check the live rate and try the model with a GPTProto API key.
by OpenAI·$1.60 in / $8.00 out per 1M tokens
Access Anthropic's Claude Opus 5.5 for coding, long-running agents, and document analysis through GPTProto. The model accepts text and images, provides a 1M-token context window and up to 128K output tokens, and uses adaptive thinking. Check the live pricing panel for GPTProto's current rates.
by Claude·$3.60 in / $18.00 out per 1M tokens
Use TypeSafe AI's Jev Latest for bounded decisions inside applications and agents. It evaluates text or structured state against questions you define, then returns typed answers and probabilities for routing, scoring, and validation. Check the GPTProto pricing panel and API Usage tab for the live rate and request forma…
by TypesafeAI·$0.034 in per 1M tokens
Access Grok 4.7 through GPTProto for coding, knowledge work, image understanding, and long-running agents. Use a 500K-token context window, configurable reasoning, and one shared API key and balance across 200+ AI models.
by Grok·$1.20 in / $3.60 out per 1M tokens
MiniMax H3 is an open-weight multimodal video model for text-to-video, first- and last-frame image-to-video, reference-led generation, and video editing. It accepts text, images, video, and audio, then produces 4–15-second clips at 768P or 2K with native stereo sound.
by MiniMax·$0.22 per video
Use the deepseek-flash model string to access DeepSeek V4.1 Flash for text, coding, vision, and agent workflows. One GPTProto API key and balance also give you access to 200+ models for evaluation, routing, and fallback design.
by DeepSeek·$0.30 in / $1.20 out per 1M tokens
OpenAI's most capable GPT Image model for exacting generation and editing. Sunburst prioritizes detail, subject fidelity, and controlled revisions for campaign creative, polished product imagery, and other quality-first workflows, with 5% off official token rates on GPTProto.
by OpenAI·$7.60 in / $28.50 out per 1M tokens
OpenAI's speed-focused GPT Image 2.5 model for everyday generation and precise editing. Use text or image inputs, choose quality from low through max, and create custom-size images up to 4K with 5% off official token rates on GPTProto.
by OpenAI·$7.60 in / $28.50 out per 1M tokens
Access Tencent's 770B-parameter Hy4-preview API through GPTProto for long-context coding, document analysis, and agent workflows. Use one API key and one shared balance to test it alongside 200+ AI models.
by Hunyuan·$0.79 in / $2.38 out per 1M tokens
Grok Imagine Image 2.0 is SpaceXAI's image generation and editing model. Its strongest surface is editing: 1,439 Elo on the LMArena image-edit board, second only to GPT-Image-2, with up to 5 reference images per call to hold a subject's identity across a series. Text-to-image is capable but not class-leading. Output is…
by Grok·$0.040 per image
Run OpenAI’s GPT-6 Astra through GPTProto for complex reasoning, coding, browser and computer-use agents, research, and professional work. Get 10% lower standard token rates and one shared balance across 200+ models.
by OpenAI·$8.00 in / $40.00 out per 1M tokens
Access Google's Gemini 3.8 API through GPTProto for long-horizon coding, autonomous agents, and multimodal analysis. Use a 1M-token context window, 64K maximum output, three thinking levels, and 40% lower token pricing with one balance shared across 200+ models.
by Google·$0.90 in / $4.50 out per 1M tokens
Access Anthropic’s highest-capability generally available model for long-running coding, research, and multi-stage knowledge work through one GPTProto API key and a shared balance across 200+ AI models.
by Claude·$9.00 in / $45.00 out per 1M tokens
Access Alibaba Qwen's September 2 coding and cowork upgrade through one GPTProto API key. Run complex engineering, research, and enterprise agents with a shared balance across 200+ models and token rates 10% below QwenCloud international pricing.
by Qwen·$1.80 in / $5.40 out per 1M tokens
Access Z.ai’s first natively multimodal GLM-5 model through GPTProto. Send text, images, videos, or files, retain up to 1M tokens of context, and use one API key and shared balance across 200+ supported models.
by Z-AI·$0.15 in / $0.50 out per 1M tokens
Call DeepSeek’s experimental multimodal V4 Flash model through GPTProto for screenshot-aware coding, chart analysis, visual QA, and tool-driven agent workflows. Use one API key and a shared balance across 200+ supported models.
by DeepSeek·$0.30 in / $1.20 out per 1M tokens
No model matches these filters.
| Model | Price: GPTProto | vs Official | vs OpenRouter | Context | Modalities | Stability | Action |
|---|---|---|---|---|---|---|---|
| API access planned | — | — | — | → | Try it now | ||
| $1.60 / $8.00$2.00 Cache read · per 1M tokens | −20% | −24% | 1.05M | → | Try it now | ||
| $1.80 / $9.00$0.18 Cache read · per 1M tokens | −10% | −15% | 1M | → | Try it now | ||
| $0.080 / $0.40$0.008 Cache read · per 1M tokens | −20% | −24% | 1.05M | → | Try it now | ||
| $1.60 / $8.00$0.16 Cache read · per 1M tokens | −20% | −24% | 1.05M | → | Try it now | ||
| $3.60 / $18.00$0.18 Cache read · per 1M tokens | −10% | −15% | 1M | → | Try it now | ||
| $0.034per 1M tokens | −20% | −24% | — | → | Try it now | ||
| $1.20 / $3.60$0.30 Cache read · per 1M tokens | −40% | −43% | 500K | → | Try it now | ||
| $0.22per video | −30% | — | — | → | Try it now | ||
| $0.30 / $1.20$0.006 Cache read · per 1M tokens | 0% | −5% | — | → | Try it now | ||
| $7.60 / $28.50$1.90 Cache read · per 1M tokens | −5% | −10% | — | → | Try it now | ||
| $7.60 / $28.50$1.90 Cache read · per 1M tokens | −5% | −10% | — | → | Try it now | ||
| $0.79 / $2.38$0.040 Cache read · per 1M tokens | −5% | −10% | 1.05M | → | Try it now | ||
| $0.040per image | 0% | — | — | → | Try it now | ||
| $8.00 / $40.00$0.80 Cache read · per 1M tokens | −20% | −24% | 1.05M | → | Try it now | ||
| $0.90 / $4.50$0.090 Cache read · per 1M tokens | −40% | −43% | 1.05M | → | Try it now | ||
| $9.00 / $45.00$0.23 Cache read · per 1M tokens | −10% | −15% | 1M | → | Try it now | ||
| $1.80 / $5.40$0.23 Cache read · per 1M tokens | −10% | −15% | 1M | → | Try it now | ||
| $0.15 / $0.50$0.030 Cache read · per 1M tokens | 0% | −5% | 1.05M | → | Try it now | ||
| $0.30 / $1.20$0.006 Cache read · per 1M tokens | 0% | −5% | 1.05M | → | Try it now |
Choose a workflow and usage level to see an example monthly cost on GPTProto, alongside the same usage at comparable provider rates.
Video and agent workloads burn credits faster — the recommendation adjusts automatically as you switch scenario or usage level.
All 20 vendors we carry. Open a tile for that vendor's models.
Choose a model and enter your monthly usage to compare estimated costs on GPTProto, the official provider, and OpenRouter.
You only pay for what you use. Top up once, spend it on any of 200+ models.
Get started. Great for testing the endpoint and trying new models.
The sweet spot for individual developers and small projects shipping to production.
Built for teams running production workloads. Bonus credits stack on already-discounted model pricing.