Repository-Scale Coding
Qwen3.7-max is tuned for multi-file repository refactoring and step-heavy Python or Rust work, and ranks among the top text models for coding in public LM Arena testing — backed by its extended-thinking reasoning.
curl --request POST "https://gptproto.com/v1/chat/completions" \
--header "Authorization: Bearer $GPTPROTO_API_KEY" \
--header "Content-Type: application/json" \
--data '{
"model": "qwen3.7-max",
"messages": [
{
"role": "user",
"content": "Hello"
}
]
}'Chat, coding agents & document work. Priced per 1M tokens — input, cached input and output are billed separately. GPTProto is 10% below official rates.
| Scenario | Qwen list | OpenRouter | GPTProto | You save / mo |
|---|---|---|---|---|
| Personal10M tokens / mo (4.8M cached) | $4.86 | $5.13 | $4.38 | −$0.49≈ $5.84 / yr |
| Team100M tokens / mo (48M cached) | $48.64 | $51.32 | $43.78 | −$4.86≈ $58.37 / yr |
| Business500M tokens / mo (240M cached) | $243.20 | $256.58 | $218.88 | −$24.32≈ $291.84 / yr |
The text-only reasoning capabilities, 1M context window, and access that define the Qwen3.7-max API on GPTProto.
Repository-Scale Coding
Qwen3.7-max is tuned for multi-file repository refactoring and step-heavy Python or Rust work, and ranks among the top text models for coding in public LM Arena testing — backed by its extended-thinking reasoning.
1M Token Qwen 3.7 Context
Qwen3.7-max ships a 1,000,000-token context window and up to 65,536 output tokens, enough to hold a full code repository or document stack in a single Qwen3.7-max API call, with no external retrieval.
Long-Horizon Qwen3.7-max Agents
Built for the agent era, Qwen3.7-max sustains multi-hour autonomous runs — Alibaba reports 1,000+ tool calls in a single session. Extended-thinking reasoning plans before it answers and holds course across long agentic loops.
One Key, Under Official Rates
Call the qwen3.7-max API through GPTProto on one balance across 200+ models — one key, no regional sign-up barriers. At $0.36 / $1.44 per 1M it runs roughly 80% below the official $2.50 / $7.50 rate.
Qwen3.7-max is the text-only reasoning flagship of Alibaba's Qwen 3.7 generation, announced at the Alibaba Cloud Summit on May 20, 2026. It is a proprietary, closed-weight model — there is no open-source qwen3.7-max release — currently available as a preview through the qwen3.7-max API. Vision and image input live in its sibling model, Qwen3.7-Plus; Max takes text in and returns text out. On GPTProto, you reach the qwen3.7-max API on the same balance and key as every other model on the platform.
| Spec | qwen3.7-max |
|---|---|
| Provider | Alibaba Cloud (Qwen) |
| Modality | Text in, text out only (no image or video input) |
| Context window | 1,000,000 tokens |
| Max output | 65,536 tokens |
| Reasoning | Extended-thinking (chain-of-thought) |
| Weights | Proprietary, closed-weight (not open source) |
| Status | Preview — announced May 20, 2026 |
| API compatibility | OpenAI-compatible; Anthropic-API compatible (works with Claude Code) |
| Model string | qwen3.7-max |
| GPTProto endpoint | https://gptproto.com/v1/chat/completions |
All four run on GPTProto under one key and one balance, so you can route per task. GPT-5.5 and Opus 4.8 still lead the Artificial Analysis Intelligence Index; qwen3.7-max is the #1 Chinese model, trailing them narrowly while undercutting both heavily on price. Against Qwen Turbo — same 1M context, same text-only modality — qwen3.7-max trades speed for far deeper extended-thinking reasoning.
| Model | Modality | Context | Direct list (in/out per 1M) | On GPTProto |
|---|---|---|---|---|
| Qwen3.7-max | Text only | 1M | $2.50 / $7.50 | $0.36 / $1.44 |
| GPT-5.5 | Multimodal | ~1.05M | $5 / $30 | $4 / $24 |
| Claude Opus 4.8 | Multimodal | 1.0M | $5 / $25 | $4 / $20 |
| Qwen Turbo | Text only | 1M | $0.05 / $0.20 | $0.045 / $0.18 |
Honest take: Reach for qwen3.7-max when you want frontier-class text reasoning, a 1M context, and long agent runs at a fraction of the cost — on GPTProto roughly 80% below its official rate. For image, diagram, or video input, use a multimodal model like GPT-5.5, Opus 4.8, or qwen3.7-max's own sibling Qwen3.7-Plus.
Moving from Alibaba Cloud Model Studio or OpenRouter to GPTProto is a drop-in change: keep your existing OpenAI-compatible request and swap two things — point base_url to https://gptproto.com/v1 and use your GPTProto API key. The same key and balance then cover qwen3.7-max and every other model on the platform, with no separate regional sign-up. Because Qwen3.7-max is also Anthropic-API compatible, Claude Code users can target it the same way.
Common questions about qwen 3.7 performance, pricing, and integration steps.
Guides, comparisons, and updates related to this model.
All Articles
GLM 5.2 is Z.ai's open-weight, MIT-licensed coding model with a 1M-token context. See its features, benchmarks vs Claude Opus 4.8 and GPT-5.5, pricing, and how to run it.

Most guides say Wan 2.7 is open source. The first-party evidence says otherwise. What Alibaba's April 2026 model actually ships — and how to run it.

When's Seedance 2.5 coming, and what will it do? We sort the confirmed facts from the leaks—and compare it to the Seedance 2.0 you can use today.

No Chinese account needed: call Seedream, Seedance and Seed from one endpoint. Which Doubao model to use, real pricing, and copy-paste code.