100万トークンのロスレスコンテキスト
GLM-5.2は、1,000,000トークンの全ウィンドウにわたって高い検索精度を維持します。これは従来のGLM-5の上限の約5倍に相当し、モノレポ全体や長大な仕様書を1回のGLM-5.2 API呼び出しで読み込めます。
curl --request POST "https://gptproto.com/v1/chat/completions" \
--header "Authorization: Bearer $GPTPROTO_API_KEY" \
--header "Content-Type: application/json" \
--data '{
"model": "glm-5.2",
"messages": [
{
"role": "user",
"content": "Hello"
}
]
}'Chat, coding agents & document work. Priced per 1M tokens — input, cached input and output are billed separately. GPTProto is 10% below official rates.
Z-AI · ≈ 148M tokens/mo (48M cached)
OpenRouter costs include its ~5.5% credit purchase fee. GPTProto applies a per-model discount (10–30% off) and your bonus credits are also spent at discounted rates — savings compound. Estimates assume a 60% cache hit rate.
GLM-5.2 APIを際立たせる要素 — 100万トークンのロスレスコンテキスト、IndexShareスパースアテンション、MITライセンスのオープンウェイト、選択可能なHigh/Max推論強度。
GLM-5.2は、1,000,000トークンの全ウィンドウにわたって高い検索精度を維持します。これは従来のGLM-5の上限の約5倍に相当し、モノレポ全体や長大な仕様書を1回のGLM-5.2 API呼び出しで読み込めます。
GLM-5.2は、長期的なコーディングタスクで数か月にわたる強化学習を行っているため、従来のGLM-5リリースで見られた早期の停滞なしに、数百回の反復にわたる自律的な複数ステップのエンジニアリング作業を継続できます。
4層ごとに1つのインデクサーを再利用するスパースアテンション設計により、1Mコンテキスト全体でトークンごとのアテンション計算量を最大2.9倍削減します。さらにマルチトークン予測を組み合わせることで、GLM-5.2 APIの推論を高速化します。
GLM-5.2の重みは、OSI承認のMITライセンスに基づくオープンソースとして提供されます。席数ごとの料金なしで、自由にセルフホスト、ファインチューニング、商用利用が可能であり、この性能クラスでは珍しい選択肢です。
GLM-5.2 is the June 2026 flagship in Z.ai's (formerly Zhipu AI) GLM-5 series — an open-weight Mixture-of-Experts model built specifically for long-horizon coding agents rather than general chat. It pairs a usable 1M-token context with native tool calling, MCP support, and structured JSON output, and ships under an MIT open-source license you can self-host. On GPTProto, the GLM-5.2 API is metered pay-as-you-go access with no regional sign-up barriers and one balance shared across 200+ models.
| Spec | GLM-5.2 |
|---|---|
| Developer | Z.ai (formerly Zhipu AI) |
| Architecture | Mixture-of-Experts, ~753B total / ~40B active per token |
| Context window | 1,000,000 tokens (lossless) |
| Max output | 131,072 tokens (~128K) |
| Modality | Text in, text out |
| Reasoning | Selectable effort — High / Max |
| Tool use | Native function calling, MCP, structured JSON |
| License | MIT (open weights) |
| Released | June 13, 2026 |
| GPTProto price | $1.26 in / $3.96 out per 1M tokens (10% under Z.ai list) |
| API model string | glm-5.2 (provider slug z-ai, full 1M context) |
All three GLM-5 releases are live on GPTProto, so moving up a version is just a model-string swap — same key, same balance. The jump to 5.2 is mostly about context size and long-task stability:
| GLM-5 | GLM-5.1 | GLM-5.2 | |
|---|---|---|---|
| Released | Feb 2026 | Apr 2026 | Jun 2026 |
| Context window | ~200K-class | ~200K-class | 1,000,000 (≈5x) |
| SWE-bench Pro* | — | 58.4 | 62.1 |
| Long-context attention | standard | standard | IndexShare (−2.9x compute) |
| Speculative decoding | — | — | MTP (+~20% accepted tokens) |
| GPTProto output / 1M | $2.88 | $3.96 | $3.96 |
The practical takeaway: GLM-5.2 lists at the same $3.96 output rate as GLM-5.1 on GPTProto but gives you roughly 5x the context and a measurable bump in long-horizon coding — so for repository-scale work there's little reason to stay on 5.1.
GPT-5.5 is also on GPTProto, so this is a real swap, not a marketing comparison. On Z.ai's reported long-horizon coding benchmarks, GLM-5.2 edges GPT-5.5 on most agentic tasks while costing far less per token:
| Benchmark (vendor-reported, Z.ai) | GLM-5.2 | GPT-5.5 |
|---|---|---|
| SWE-bench Pro | 62.1 | 58.6 |
| FrontierSWE (long-horizon) | 74.4% | 72.6% |
| MCP-Atlas (tool use) | 77.0 | 75.3 |
| Humanity's Last Exam (w/ tools) | 54.7 | 52.2 |
| Terminal-Bench 2.1 | 81.0 | 84.0 |
| GPTProto price / 1M (in / out) | $1.26 / $3.96 | $4 / $24 |
On GPTProto the GLM-5.2 API runs about 3x cheaper on input and 6x cheaper on output than GPT-5.5 ($1.26/$3.96 vs $4/$24 per 1M). GPT-5.5 still leads on raw terminal tasks, so it isn't a clean sweep — but for agentic, tool-driven, multi-step coding the benchmark edge favors GLM-5.2 at a fraction of the spend. On the independent Artificial Analysis Intelligence Index v4.1, GLM-5.2 scores 51, competitive at the open frontier though not ahead of every closed model on general tasks.
If you're already calling GLM-5.2 on Z.ai's API, moving to GPTProto is a drop-in: keep your request body, change the base URL and key. You get the same glm-5.2 model with no separate Z.ai account, no regional sign-up friction, and one balance that also covers GPT-5.5, Claude, Gemini, and 200+ other models. The exact endpoint and model string are in the Quick Start below.
Zhipu AIのオープンウェイトフレームワークによるGLM 5.2の性能、コンテキスト制限、ローカルデプロイについての疑問にお答えします。
このモデルに関連するガイド、比較、最新情報。
すべての記事
GLM 5.2は、Z.aiが提供するMITライセンスのオープンウェイト・コーディングモデルで、100万トークンのコンテキストに対応しています。機能、Claude Opus 4.8およびGPT-5.5とのベンチマーク比較、料金、実行方法をご紹介します。

2026年におすすめのText-to-Image API 7選を、Arena Eloと実際の料金でランキング。GPT Image 2、Nano Banana、Seedream 5.0などを、1つのAPIキーですべて利用できます。

Nano Banana 2はNano Banana Proの半額で、Image Arenaではより高いスコアを獲得しています。どちらのGemini画像モデルが優れているのか、用途別に解説。1つのAPIで両方を実行できるコードも紹介します。

印刷対応(300 DPI)でキャラクターの一貫性を保ったAI絵本を構築できます。すべてのページに同じキャラクターを登場させ、KDP対応ファイルを、APIコスト約1ドルで作成。完全なコード付き。