GPT Proto

GPTProto

  • 儀表板
  • LLM

    • google
      Gemini 3.7 Flash新功能
    • grok
      Grok 4.6
    • qwen
      Qwen3.8 Max
    • claude
      Claude Opus 5
    • google
      Gemini 3.6 Flash

    影像

    • bytedance
      Dola Seedream 5.0 Pro 260628新功能
    • google
      Gemini 3.1 Flash Lite Image
    • google
      Gemini 3.1 Flash Image
    • openai
      GPT Image 2
    • google
      Gemini 3.1 Flash Image Preview

    影片

    • bytedance
      Dreamina Seedance 2.5 260628新功能
    • kling
      Kling v3.0 4k
    • bytedance
      Dreamina Seedance 2.0 Mini 260615
    • kling
      Kling v3 Omni 4k
    • bytedance
      Dreamina Seedance 2.0 Fast 260128
    探索 218+ 種模型 >
  • 生成器

    • 建立圖片
    • 建立影片
    • 在畫布中編輯

    功能

    • 動漫轉真人 AI新功能
    • 動漫 AI 藝術生成器
    • AI 物件移除器
    • AI 圖片編輯器
    • 無限制 AI 圖像生成器
    • AI 動作轉移
    • AI 衣物移除器
    • AI 浮水印移除工具
    • 線上 AI 圖片增強器
    • 線上背景移除工具
    探索全部 >

    提示詞

    • Seedance 2.0 提示詞新功能
    • GPT Image 2 提示詞
    • Nano Banana Pro 提示詞
    • Seedream 5.0 Pro 提示詞
  • AI 部落格

    • Seedance 2.5 與 MiniMax H3:哪個更適合製作電商廣告?
    • GLM 5.2 與 Claude Opus 5:哪個程式編碼模型更具成本效益?
    • GLM 5.2 與 MiniMax M3:哪個更適合程式設計與前端工作?
    • 如何使用 API 建立自己的 AI 角色——無需撰寫程式碼
    • Kimi K3 與 Claude Opus 5:哪個更適合程式設計與 AI 代理?
    探索全部 >

    AI 洞察

    • MiniMax H3 正式登場:其影片編輯升級實際帶來哪些改變
    • 什麼是 Emochi AI?以及它為什麼成長得這麼快?(2026)
    • Kimi K3 是什麼?真的接近 GPT-5.6 與 Fable 5 嗎?
    • 2026 年 YouTube、TikTok、文字與圖片適用的 12 款最佳 AI 影片生成工具
    • Gemini 3.6 Flash 與 Gemini 3.5 Flash-Lite 詳解:您應該使用哪一個?
    探索全部 >

    AI 文件

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    探索全部 >

    AI 技能

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    探索全部 >
定價+7% 贈送
English繁體中文한국어日本語EspañolРусский
立即開始
  1. 首頁
  2. /模型
  3. /Google
  4. /gemini-3.7-flash
對話
Google
Gemini 3.7 Flash
$ 
對話
對話
Google Gemini 3.7 Flash is a GA multimodal reasoning model for coding, agents, UI work, and document analysis. GPTProto provides Gemini 3.7 Flash API access at 40% off through one API key and a shared balance across 200+ models.

模態

輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字

/

Gemini 3.7 Flash pricing

Estimate a request with real work scenarios. GPTProto token pricing is 40% below official rates.

Cost calculator

Multi-turn agent with cached context.
TokensRateCost
$0.45 / 1M$0.000675
$2.25 / 1M$0.0018
$0.045 / 1M$0.001125
Cost per request$0.0036

Top up

GPTProto vs official pricing.
Requests
You pay
40% off
$100
You receive$100.00

Save$66.66 (40%)vs Google

Gemini 3.7 Flash API for Coding and Agent Workflows

Run Google's newest Flash model through GPTProto with multimodal input, a 1,048,576-token context window, configurable reasoning, and OpenAI-compatible calls for coding, tool use, structured output, and long-document workflows.

1M-Token Multimodal Context

Process text, images, video, audio, and PDFs in one request. The 1,048,576-token input window supports large repositories, long reports, recorded meetings, and multi-file analysis without reducing everything to short excerpts.

Coding and UI Implementation

Generate, debug, and refactor code while using screenshots or design references as context. Gemini 3.7 Flash improves first-pass code accuracy and design adherence over Gemini 3.6 Flash on Google's published evaluations.

Agent Tools and Structured Output

Use function calling, structured outputs, code execution, file search, URL context, and search grounding where the selected API route supports them. This makes the model suitable for multi-step agents that must inspect, act, and return validated data.

Low, Medium, or High Reasoning

Choose low for faster routine work, medium for balanced coding and agents, or high for difficult reasoning and tool use. Gemini 3.7 Flash does not support the minimal thinking level.

相關模型
所有模型
模型輸入 → 輸出
Gemini 3.7 Flash目前
—$0.45 / $2.25 每 1M— / $0.04 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Grok 4.6
500K$1.20 / $3.60 每 1M— / $0.30 每 1M
輸入: 文字輸入: 圖像
輸出: 文字
Qwen3.8 Max
1M$1.80 / $5.40 每 1M$2.25 / $0.23 每 1M
輸入: 文字輸入: 圖像輸入: 影片輸入: 文件
輸出: 文字
Claude Opus 5
1M$4.00 / $20.00 每 1M$5.00 / $0.40 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Gemini 3.6 Flash
1.05M$0.45 / $2.25 每 1M— / $0.04 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Gemini 3.5 Flash Lite
1.05M$0.18 / $1.50 每 1M$0.02 / $0.02 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Kimi K3
1.05M$2.70 / $13.50 每 1M$0.27 / $0.27 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
GPT 5.6 Luna
1.05M$0.16 / $0.96 每 1M$0.20 / $0.02 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
GPT 5.6 Terra
1.05M$1.60 / $9.60 每 1M$2.00 / $0.16 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
GPT 5.6 Sol
1.05M$4.00 / $24.00 每 1M$5.00 / $0.40 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Grok 4.5
500K$1.20 / $3.60 每 1M$0.30 / $0.30 每 1M
輸入: 文字輸入: 圖像
輸出: 文字
Claude Sonnet 5
1M$1.60 / $8.00 每 1M$2.00 / $0.16 每 1M
輸入: 文字輸入: 文件
輸出: 文字
Minimax M3
1.05M$0.48 / $0.96 每 1M$0.10 / $0.10 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
GLM 5.2
1.05M$1.26 / $3.96 每 1M$0.23 / $0.23 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Claude Fable 5
1M$8.00 / $40.00 每 1M$10.00 / $0.80 每 1M
輸入: 文字輸入: 文件
輸出: 文字
Qwen3.7 Max
1M$0.36 / $1.44 每 1M$0.07 / $0.07 每 1M
輸入: 文字輸入: 文件
輸出: 文字
Gemini 3.5 Flash
1.05M$0.90 / $5.40 每 1M$0.09 / $0.09 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
DeepSeek v4 Flash
—1.05M$0.14 / $0.28 每 1M— / $0.0028 每 1M
輸入: 文字
輸出: 文字
DeepSeek v4 Pro
1.05M$1.04 / $2.09 每 1M$0.0087 / $0.0087 每 1M
輸入: 文字
輸出: 文字
Grok 4.3
1M$0.75 / $1.50 每 1M$0.12 / $0.12 每 1M
輸入: 文字輸入: 圖像
輸出: 文字
Kimi K2.6
262K$0.85 / $3.60 每 1M$0.14 / $0.14 每 1M
輸入: 文字輸入: 文件
輸出: 文字
GLM 5.1
205K$1.26 / $3.96 每 1M$0.23 / $0.23 每 1M
輸入: 文字輸入: 文件
輸出: 文字
GLM 5 Turbo
203K$1.08 / $3.60 每 1M$0.22 / $0.22 每 1M
輸入: 文字輸入: 文件
輸出: 文字
Gemini 3.1 Flash Lite Preview
1.05M$0.15 / $0.90 每 1M$0.01 / $0.01 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
DeepSeek v3.2
164K$0.17 / $0.25 每 1M$0.02 / $0.02 每 1M
輸入: 文字
輸出: 文字
Minimax M2.5
205K$0.24 / $0.96 每 1M$0.30 / $0.02 每 1M
輸入: 文字輸入: 文件
輸出: 文字
Gemini 3.1 Pro Preview
1.05M$1.20 / $7.20 每 1M$0.12 / $0.12 每 1M
輸入: 文字輸入: 圖像輸入: 文件
輸出: 文字
Kimi K2.5
262K$0.54 / $2.70 每 1M$0.09 / $0.09 每 1M
輸入: 文字輸入: 文件
輸出: 文字
Qwen Turbo
—$0.04 / $0.18 每 1M$0.009 / $0.009 每 1M
輸入: 文字
輸出: 文字
Gemini 3 Flash Preview
1.05M$0.30 / $1.80 每 1M$0.03 / $0.03 每 1M
輸入: 文字輸入: 圖像
輸出: 文字
Doubao Seed 1.6 Thinking 250715
262K$0.10 / $0.97 每 1M—
輸入: 文字輸入: 圖像
輸出: 文字
Doubao Seed 1.6 Thinking 250615
262K$0.10 / $0.97 每 1M—
輸入: 文字輸入: 圖像
輸出: 文字
Doubao Seed 1.6 Flash 250615
262K$0.02 / $0.18 每 1M—
輸入: 文字輸入: 圖像
輸出: 文字

What Is the Gemini 3.7 Flash API?

The Gemini 3.7 Flash API gives developers programmatic access to Google's newest generally available Flash model. Released on August 13, 2026, it is designed for software engineering, web development, multimodal document work, and multi-step agents that need stronger instruction following and more reliable tool use than the previous Flash release.

Gemini 3.7 Flash accepts text, images, video, audio, and PDF input, but it returns text only. It is therefore suitable for understanding a product screenshot, reviewing a recorded workflow, extracting evidence from reports, or reasoning across a repository. It is not a native image, video, or audio generation model.

On GPTProto, developers can access the model without maintaining a separate balance for every provider. The same key can route requests across 200+ supported models, making it easier to compare Gemini with alternatives or add a fallback without opening another provider account. The live pricing panel is the source of truth for current Gemini 3.7 Flash API pricing and the displayed 40% discount.

Specification Gemini 3.7 Flash
Provider Google
Release status Generally available (GA)
Release date August 13, 2026
Official model ID gemini-3.7-flash
Input modalities Text, image, video, audio, and PDF
Output modality Text
Input token limit 1,048,576 tokens
Maximum output 65,536 tokens
Thinking levels low, medium (default), high
Supported upstream capabilities Caching, function calling, structured outputs, code execution, file search, URL context, search and Maps grounding, Computer Use (preview)
Not supported upstream minimal thinking, native image generation, audio generation, and Live API
Consumption options upstream Standard, Batch, Flex, and Priority

Gemini 3.7 Flash API Applications

The Gemini 3.7 Flash API is best suited to workloads that combine several steps instead of asking for a single short answer. For coding, it can inspect repository context, trace an issue across files, propose a minimal patch, and explain the change. For frontend work, a screenshot or design reference can be included with the request so the model can compare the intended layout with the implementation.

For agent workflows, the model can decide when to call a function, interpret tool results, and return a response that follows a JSON schema. Practical uses include support-ticket triage, document-to-database extraction, compliance review, internal research, and workflow automation. Developers can lower reasoning for time-sensitive classification or raise it for difficult debugging and planning tasks.

Its multimodal input is also useful for knowledge work. A single workflow can combine PDFs, charts, screenshots, recorded meetings, and written instructions. Because output remains text, use a dedicated generation model when the final deliverable must be an image, video, or audio file.

Gemini 3.7 Flash vs Gemini 3.6 Flash

Gemini 3.7 Flash keeps the same 1,048,576-token input limit, 65,536-token output limit, and broad input modalities as Gemini 3.6 Flash. The reason to upgrade is not a larger context window. It is the improvement in coding, web development, document comprehension, instruction following, and agent execution.

Google's launch evaluations report the following changes. These are provider-reported benchmark results, not independent GPTProto tests.

Evaluation Gemini 3.7 Flash Gemini 3.6 Flash Reported change
FrontierCode 1.1 Main 43.6% 34.4% +9.2 points
DeepSWE v1.1 65.3% 49.0% +16.3 points
WebDev Arena 1588 Elo 1538 Elo +50 Elo
GDP.pdf 34.0% 22.0% +12.0 points
AutomationBench 30.4% 17.0% +13.4 points

There is one compatibility tradeoff: Gemini 3.6 Flash accepts minimal, low, medium, and high thinking levels, while 3.7 supports only low, medium, and high. If an existing workflow relies on near-zero thinking for simple, high-volume requests, keep 3.6 in an A/B test or compare it with a Flash-Lite model before migrating all traffic.

Migration Details to Check Before You Switch

Changing the model name is only the first step. Use the following checklist before replacing Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3 Flash Preview, or Gemini 3.1 Pro in an existing application:

  • Use the exact model string exposed in GPTProto documentation. Google's official stable ID is gemini-3.7-flash; confirm that the GPTProto route uses the same string before publishing or deploying.

  • Do not send thinking_level: "minimal". Google documents medium as the default and returns a validation error when minimal is selected.

  • For native Gemini Interactions API migrations, remove deprecated temperature, top_p, top_k, and candidate_count settings, replace thinking_budget with thinking_level, and remove prefilled model turns.

  • Keep multimodal expectations precise: images, audio, video, and PDFs are input formats, while the response is text.

  • Do not assume every upstream Google tool is automatically available through an OpenAI-compatible route. Check the live GPTProto API documentation for supported parameters, tools, file handling, and response fields.

  • Run a canary test with your own coding tasks, schemas, tool calls, latency targets, and token budgets before moving all requests.

This migration guidance is a major difference between a basic model listing and a usable Gemini 3.7 Flash API access page: it tells developers which existing requests can fail even when the new model ID is correct.

When Should You Choose Gemini 3.7 Flash?

Choose Gemini 3.7 Flash when a workload needs a long context window, multiple input formats, stronger first-pass coding, or multi-step tool orchestration at a lower token price than larger flagship models. Keep Gemini 3.6 Flash when minimal thinking is important or when an existing application has not yet passed its regression tests.

For Grok 4.6 vs Gemini 3.7 Flash, the decision is more specific. Gemini accepts audio, video, and PDF input and offers more than twice Grok's context window. Grok 4.6 adds an xhigh reasoning option and SpaceXAI documents no fixed text-output limit. Neither model is universally better, so route representative tasks through both before selecting a default.

Decision factor Gemini 3.7 Flash Gemini 3.6 Flash Grok 4.6
Best fit Coding, design adherence, multimodal documents, multi-step agents Existing Flash workflows and minimal-thinking traffic Coding and agents needing xhigh reasoning or very long text output
Context window 1,048,576 tokens 1,048,576 tokens 500,000 tokens
Native inputs Text, image, video, audio, PDF Text, image, video, audio, PDF Text, image
Text output limit 65,536 tokens 65,536 tokens No fixed limit documented by SpaceXAI
Reasoning control Low, medium, high Minimal, low, medium, high Low, medium, high, xhigh
Official short-context token price* $0.75 input / $3.75 output through Dec. 31, 2026 $0.75 input / $3.75 output through Dec. 31, 2026 $2 input / $6 output below 200K input tokens

*Official provider pricing is shown only for model selection context. Use GPTProto's live pricing panel for the rate billed through GPTProto.

Gemini 3.7 Flash API: Common Technical Questions

How much does the Gemini 3.7 Flash API cost on GPTProto?

GPTProto displays a 40% off rate in the live pricing panel. Use that panel for the current input and output prices because provider promotions and routing rates can change. Do not use Google's native API price as a substitute for the GPTProto billing rate.

How can I get a Gemini 3.7 Flash API key?

Create a GPTProto account, add balance, and generate one API key from the dashboard. The same key and balance can be used across Gemini 3.7 Flash and 200+ other supported models, so you do not need a separate credential for each provider.

What is the Gemini 3.7 Flash context window?

The model accepts up to 1,048,576 input tokens and can return up to 65,536 output tokens. A large advertised window does not guarantee that every retrieval or reasoning task will be accurate at the limit, so test with documents and repository structures that match your real workload.

Does Gemini 3.7 Flash support image, video, audio, and PDF input?

Yes. It accepts text, images, video, audio, and PDFs as input and returns text. It does not natively generate images or audio, and it does not support Google's Live API.

Is Gemini 3.7 Flash good for coding and AI agents?

Yes. Google positions it for coding and agents, and its launch results show gains over 3.6 Flash on FrontierCode, DeepSWE, WebDev Arena, and AutomationBench. Validate it on your own repositories, tool definitions, and acceptance tests before replacing a production model.

What changed from Gemini 3.6 Flash to Gemini 3.7 Flash?

The context and output limits remain the same, but 3.7 improves coding, design adherence, document reasoning, and multi-step execution in Google's evaluations. The main configuration difference is that 3.7 removes the minimal thinking level.

When was the Gemini 3.7 Flash API released?

Google released Gemini 3.7 Flash as a generally available model on August 13, 2026. The stable official model ID is gemini-3.7-flash, with no preview suffix.

Can I use Google's native Gemini SDK examples with a GPTProto key?

Not without changing the integration. Native Google SDK examples use Google's authentication, endpoint, and request schema. GPTProto uses its own documented, OpenAI-compatible access route, so follow the live GPTProto Quick Start and parameter reference.

Which is better: Grok 4.6 or Gemini 3.7 Flash?

Gemini 3.7 Flash is the clearer fit for 1M-context work and audio, video, or PDF understanding. Grok 4.6 is worth testing when xhigh reasoning or unrestricted text-output length matters. For coding quality, compare both on the same repository tasks rather than choosing from one benchmark or token price alone.

相關文章

與本模型相關的指南、對比與更新。

所有文章
What Is GLM-5.3? Z.ai's Quiet Coding Plan Launch, Pricing, and Confirmed Upgrades

What Is GLM-5.3? Z.ai's Quiet Coding Plan Launch, Pricing, and Confirmed Upgrades

GLM-5.3 is live in Z.ai’s Coding Plan. See its release status, 1M context, reasoning modes, pricing, upgrades, and remaining unknowns.

7 Best Image Editing AI Models in 2026 for API, Batch Editing, and Product Photos

7 Best Image Editing AI Models in 2026 for API, Batch Editing, and Product Photos

Compare 7 image editing AI models for APIs, batch workflows, product photos, text edits, and brand consistency, with the best pick for each task.

DeepSeek V4 Pro vs Kimi K3: What Changed After the 0813 Update?

DeepSeek V4 Pro vs Kimi K3: What Changed After the 0813 Update?

Compare DeepSeek V4 Pro 0813 vs Kimi K3 for coding, speed, multimodal input, and API cost—and see which model better fits your project.

Grok 4.6 vs DeepSeek V4 Pro: Coding, Pricing, and Which Is Better?

Grok 4.6 vs DeepSeek V4 Pro: Coding, Pricing, and Which Is Better?

Grok 4.6 vs DeepSeek V4 Pro compared on coding, frontend work, benchmarks, context and API pricing. See which model offers better value for developers.

GPT Proto

以全球規模與穩定性,賦能 AI 創新:

透過我們的旗艦產品 GPT Proto,我們提供統一的介面,讓您能存取並整合全球頂尖 AI 供應商的 API,涵蓋文字、視覺、語音等領域。我們協助開發者與企業簡化整合流程,並無限制地加速創新。

全球基礎設施,在地合規:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

為擴展而生:

我們深知穩定性至關重要。我們的平台建立在強大的去中心化架構之上,支援動態自動擴展。無論您是進行試點專案還是處理數百萬次併發請求,我們的系統都能即時擴展以滿足需求,確保您的業務永遠不會受限於基礎設施。

導覽

  • 儀表板
  • 模型
  • 建立圖片
  • AI 圖片放大
  • AI 背景移除
  • 建立影片
  • 在畫布中編輯
  • 功能
  • 定價
  • AI 文件
  • AI 部落格
  • AI 洞察
  • AI 技能

功能

  • 動漫轉真人 AI
  • 動漫 AI 藝術生成器
  • AI 物件移除器
  • AI 圖片編輯器
  • 無限制 AI 圖像生成器
  • AI 動作轉移
  • AI 衣物移除器
  • AI 浮水印移除工具
  • 線上 AI 圖片增強器
  • 線上背景移除工具
  • AI 臉部交換圖片
  • AI 護照照片製作器
  • MS Paint AI 生成器
Explore all features >

LLM

  • Gemini 3.7 Flash
  • Grok 4.6
  • Qwen3.8 Max
  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
  • GPT 5.6 Sol
  • Grok 4.5
  • Claude Sonnet 5
  • Minimax M3
  • GLM 5.2
  • GPT 5.1 Chat Latest
  • Claude Fable 5
  • Qwen3.7 Max
  • Claude Opus 4.8 Thinking
  • Claude Opus 4.8
  • Gemini 3.5 Flash
探索所有模型 >

影像

  • Dola Seedream 5.0 Pro 260628
  • Gemini 3.1 Flash Lite Image
  • Gemini 3.1 Flash Image
  • GPT Image 2
  • Gemini 3.1 Flash Image Preview
  • Seedream 5.0 260128
  • Doubao Seedream 5.0 260128
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image O1
  • GPT Image 1.5
  • Seedream 4.5 251128
  • Doubao Seedream 4.5 251128
  • Grok Imagine 0.9
  • Gemini 3 Pro Image Preview
  • Qwen Image Lora
  • Qwen Image Plus Lora
  • Qwen Image Plus
  • Grok 4 Image
  • GPT Image 1 Mini
探索所有模型 >

影片

  • Dreamina Seedance 2.5 260628
  • Kling v3.0 4k
  • Dreamina Seedance 2.0 Mini 260615
  • Kling v3 Omni 4k
  • Dreamina Seedance 2.0 Fast 260128
  • Dreamina Seedance 2.0 260128
  • Vidu 2.0
  • Doubao Seedance 2.0 260128
  • Doubao Seedance 2.0 Fast 260128
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Vidu Q3 Turbo
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
  • Vidu Q2 Pro Fast
  • Vidu Q2
探索所有模型 >

© 2026 Talent Tech Global Limited (Hong Kong). 保留所有權利。

註冊地址: Unit 1022a, Beverley Commercial Centre, 87-105 Chatham Road South, Tsim Sha Tsui, Hong Kong商業登記證號碼: 79462435-000-12-25-0
  • 關於我們
  • 隱私權政策
  • 服務條款
  • 網站地圖