GPT Proto

GPTProto

  • Dashboard
  • LLM

    • google
      Gemini 3.7 FlashNew
    • grok
      Grok 4.6
    • qwen
      Qwen3.8 Max
    • claude
      Claude Opus 5
    • google
      Gemini 3.6 Flash

    Image

    • bytedance
      Dola Seedream 5.0 Pro 260628New
    • google
      Gemini 3.1 Flash Lite Image
    • google
      Gemini 3.1 Flash Image
    • openai
      GPT Image 2
    • google
      Gemini 3.1 Flash Image Preview

    Video

    • bytedance
      Dreamina Seedance 2.5 260628New
    • kling
      Kling v3.0 4k
    • bytedance
      Dreamina Seedance 2.0 Mini 260615
    • kling
      Kling v3 Omni 4k
    • bytedance
      Dreamina Seedance 2.0 Fast 260128
    Explore 218+ Models >
  • Generator

    • Create Image
    • Create Video
    • Edit in Canvas

    Features

    • AI Packaging Design GeneratorNew
    • Anime to Real Life AI
    • Anime AI Art Generator
    • AI Object Remover
    • AI Image Editor
    • AI Motion Transfer
    • AI Watermark Remover
    • AI Image Enhancer Online
    • Online Background Remover Tool
    • AI Face Swap Image
    Explore All >

    Prompts

    • Seedance 2.0 PromptsNew
    • GPT Image 2 Prompts
    • Nano Banana Pro Prompts
    • Seedream 5.0 Pro Prompts
  • AI Blog

    • 7 Best Image Editing AI Models in 2026 for API, Batch Editing, and Product Photos
    • DeepSeek V4 Pro vs Kimi K3: What Changed After the 0813 Update?
    • Grok 4.6 vs DeepSeek V4 Pro: Coding, Pricing, and Which Is Better?
    • Grok 4.6 vs Kimi K3: Which One Fits Your Project?
    • Seedance 2.0 vs Seedance 2.5: Same Prompt, Storyboard, and Real Results
    Explore All >

    AI Insight

    • What Is GLM-5.3? Z.ai's Quiet Coding Plan Launch, Pricing, and Confirmed Upgrades
    • What Is OpenAI's Newest Model Astra? Release Date, Benchmarks & How It Compares (2026)
    • MiniMax H3 Is Here: What Its Video Editing Upgrade Actually Changes
    • What Is Emochi AI—and Why Is It Growing So Fast? (2026)
    • What Is Kimi K3—and Is It Really Close to GPT-5.6 and Fable 5?
    Explore All >

    AI Docs

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    Explore All >

    AI Skills

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    Explore All >
Pricing+7% bonus
English繁體中文한국어日本語EspañolРусский
Get Started Now
  1. Home
  2. /Model
  3. /Google
  4. /gemini-3.7-flash
Chat
Google
Gemini 3.7 Flash
$ 
Chat
Chat
Google Gemini 3.7 Flash is a GA multimodal reasoning model for coding, agents, UI work, and document analysis. GPTProto provides Gemini 3.7 Flash API access at 40% off through one API key and a shared balance across 200+ models.

Modalities

Input: TextInput: ImageInput: Document
Output: Text

/

Gemini 3.7 Flash pricing

Estimate a request with real work scenarios. GPTProto token pricing is 40% below official rates.

Cost calculator

Multi-turn agent with cached context.
TokensRateCost
$0.45 / 1M$0.000675
$2.25 / 1M$0.0018
$0.045 / 1M$0.001125
Cost per request$0.0036

Top up

GPTProto vs official pricing.
Requests
You pay
40% off
$100
You receive$100.00

Save$66.66 (40%)vs Google official

Gemini 3.7 Flash API for Coding and Agent Workflows

Run Google's newest Flash model through GPTProto with multimodal input, a 1,048,576-token context window, configurable reasoning, and OpenAI-compatible calls for coding, tool use, structured output, and long-document workflows.

1M-Token Multimodal Context

Process text, images, video, audio, and PDFs in one request. The 1,048,576-token input window supports large repositories, long reports, recorded meetings, and multi-file analysis without reducing everything to short excerpts.

Coding and UI Implementation

Generate, debug, and refactor code while using screenshots or design references as context. Gemini 3.7 Flash improves first-pass code accuracy and design adherence over Gemini 3.6 Flash on Google's published evaluations.

Agent Tools and Structured Output

Use function calling, structured outputs, code execution, file search, URL context, and search grounding where the selected API route supports them. This makes the model suitable for multi-step agents that must inspect, act, and return validated data.

Low, Medium, or High Reasoning

Choose low for faster routine work, medium for balanced coding and agents, or high for difficult reasoning and tool use. Gemini 3.7 Flash does not support the minimal thinking level.

Related Models
All Models
ModelInput → Output
Gemini 3.7 FlashCurrent
—$0.45 / $2.25 per 1M— / $0.04 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Grok 4.6
500K$1.20 / $3.60 per 1M— / $0.30 per 1M
Input: TextInput: Image
Output: Text
Qwen3.8 Max
1M$1.80 / $5.40 per 1M$2.25 / $0.23 per 1M
Input: TextInput: ImageInput: VideoInput: Document
Output: Text
Claude Opus 5
1M$4.00 / $20.00 per 1M$5.00 / $0.40 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Gemini 3.6 Flash
1.05M$0.45 / $2.25 per 1M— / $0.04 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Gemini 3.5 Flash Lite
1.05M$0.18 / $1.50 per 1M$0.02 / $0.02 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Kimi K3
1.05M$2.70 / $13.50 per 1M$0.27 / $0.27 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GPT 5.6 Luna
1.05M$0.16 / $0.96 per 1M$0.20 / $0.02 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GPT 5.6 Terra
1.05M$1.60 / $9.60 per 1M$2.00 / $0.16 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GPT 5.6 Sol
1.05M$4.00 / $24.00 per 1M$5.00 / $0.40 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Grok 4.5
500K$1.20 / $3.60 per 1M$0.30 / $0.30 per 1M
Input: TextInput: Image
Output: Text
Claude Sonnet 5
1M$1.60 / $8.00 per 1M$2.00 / $0.16 per 1M
Input: TextInput: Document
Output: Text
Minimax M3
1.05M$0.48 / $0.96 per 1M$0.10 / $0.10 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GLM 5.2
1.05M$1.26 / $3.96 per 1M$0.23 / $0.23 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Claude Fable 5
1M$8.00 / $40.00 per 1M$10.00 / $0.80 per 1M
Input: TextInput: Document
Output: Text
Qwen3.7 Max
1M$0.36 / $1.44 per 1M$0.07 / $0.07 per 1M
Input: TextInput: Document
Output: Text
Gemini 3.5 Flash
1.05M$0.90 / $5.40 per 1M$0.09 / $0.09 per 1M
Input: TextInput: ImageInput: Document
Output: Text
DeepSeek v4 Flash
—1.05M$0.14 / $0.28 per 1M— / $0.0028 per 1M
Input: Text
Output: Text
DeepSeek v4 Pro
1.05M$1.04 / $2.09 per 1M$0.0087 / $0.0087 per 1M
Input: Text
Output: Text
Grok 4.3
1M$0.75 / $1.50 per 1M$0.12 / $0.12 per 1M
Input: TextInput: Image
Output: Text
Kimi K2.6
262K$0.85 / $3.60 per 1M$0.14 / $0.14 per 1M
Input: TextInput: Document
Output: Text
GLM 5.1
205K$1.26 / $3.96 per 1M$0.23 / $0.23 per 1M
Input: TextInput: Document
Output: Text
GLM 5 Turbo
203K$1.08 / $3.60 per 1M$0.22 / $0.22 per 1M
Input: TextInput: Document
Output: Text
Gemini 3.1 Flash Lite Preview
1.05M$0.15 / $0.90 per 1M$0.01 / $0.01 per 1M
Input: TextInput: ImageInput: Document
Output: Text
DeepSeek v3.2
164K$0.17 / $0.25 per 1M$0.02 / $0.02 per 1M
Input: Text
Output: Text
Minimax M2.5
205K$0.24 / $0.96 per 1M$0.30 / $0.02 per 1M
Input: TextInput: Document
Output: Text
Gemini 3.1 Pro Preview
1.05M$1.20 / $7.20 per 1M$0.12 / $0.12 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Kimi K2.5
262K$0.54 / $2.70 per 1M$0.09 / $0.09 per 1M
Input: TextInput: Document
Output: Text
Qwen Turbo
—$0.04 / $0.18 per 1M$0.009 / $0.009 per 1M
Input: Text
Output: Text
Gemini 3 Flash Preview
1.05M$0.30 / $1.80 per 1M$0.03 / $0.03 per 1M
Input: TextInput: Image
Output: Text
Doubao Seed 1.6 Thinking 250715
262K$0.10 / $0.97 per 1M—
Input: TextInput: Image
Output: Text
Doubao Seed 1.6 Thinking 250615
262K$0.10 / $0.97 per 1M—
Input: TextInput: Image
Output: Text
Doubao Seed 1.6 Flash 250615
262K$0.02 / $0.18 per 1M—
Input: TextInput: Image
Output: Text

What Is the Gemini 3.7 Flash API?

The Gemini 3.7 Flash API gives developers programmatic access to Google's newest generally available Flash model. Released on August 13, 2026, it is designed for software engineering, web development, multimodal document work, and multi-step agents that need stronger instruction following and more reliable tool use than the previous Flash release.

Gemini 3.7 Flash accepts text, images, video, audio, and PDF input, but it returns text only. It is therefore suitable for understanding a product screenshot, reviewing a recorded workflow, extracting evidence from reports, or reasoning across a repository. It is not a native image, video, or audio generation model.

On GPTProto, developers can access the model without maintaining a separate balance for every provider. The same key can route requests across 200+ supported models, making it easier to compare Gemini with alternatives or add a fallback without opening another provider account. The live pricing panel is the source of truth for current Gemini 3.7 Flash API pricing and the displayed 40% discount.

Specification Gemini 3.7 Flash
Provider Google
Release status Generally available (GA)
Release date August 13, 2026
Official model ID gemini-3.7-flash
Input modalities Text, image, video, audio, and PDF
Output modality Text
Input token limit 1,048,576 tokens
Maximum output 65,536 tokens
Thinking levels low, medium (default), high
Supported upstream capabilities Caching, function calling, structured outputs, code execution, file search, URL context, search and Maps grounding, Computer Use (preview)
Not supported upstream minimal thinking, native image generation, audio generation, and Live API
Consumption options upstream Standard, Batch, Flex, and Priority

Gemini 3.7 Flash API Applications

The Gemini 3.7 Flash API is best suited to workloads that combine several steps instead of asking for a single short answer. For coding, it can inspect repository context, trace an issue across files, propose a minimal patch, and explain the change. For frontend work, a screenshot or design reference can be included with the request so the model can compare the intended layout with the implementation.

For agent workflows, the model can decide when to call a function, interpret tool results, and return a response that follows a JSON schema. Practical uses include support-ticket triage, document-to-database extraction, compliance review, internal research, and workflow automation. Developers can lower reasoning for time-sensitive classification or raise it for difficult debugging and planning tasks.

Its multimodal input is also useful for knowledge work. A single workflow can combine PDFs, charts, screenshots, recorded meetings, and written instructions. Because output remains text, use a dedicated generation model when the final deliverable must be an image, video, or audio file.

Gemini 3.7 Flash vs Gemini 3.6 Flash

Gemini 3.7 Flash keeps the same 1,048,576-token input limit, 65,536-token output limit, and broad input modalities as Gemini 3.6 Flash. The reason to upgrade is not a larger context window. It is the improvement in coding, web development, document comprehension, instruction following, and agent execution.

Google's launch evaluations report the following changes. These are provider-reported benchmark results, not independent GPTProto tests.

Evaluation Gemini 3.7 Flash Gemini 3.6 Flash Reported change
FrontierCode 1.1 Main 43.6% 34.4% +9.2 points
DeepSWE v1.1 65.3% 49.0% +16.3 points
WebDev Arena 1588 Elo 1538 Elo +50 Elo
GDP.pdf 34.0% 22.0% +12.0 points
AutomationBench 30.4% 17.0% +13.4 points

There is one compatibility tradeoff: Gemini 3.6 Flash accepts minimal, low, medium, and high thinking levels, while 3.7 supports only low, medium, and high. If an existing workflow relies on near-zero thinking for simple, high-volume requests, keep 3.6 in an A/B test or compare it with a Flash-Lite model before migrating all traffic.

Migration Details to Check Before You Switch

Changing the model name is only the first step. Use the following checklist before replacing Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3 Flash Preview, or Gemini 3.1 Pro in an existing application:

  • Use the exact model string exposed in GPTProto documentation. Google's official stable ID is gemini-3.7-flash; confirm that the GPTProto route uses the same string before publishing or deploying.

  • Do not send thinking_level: "minimal". Google documents medium as the default and returns a validation error when minimal is selected.

  • For native Gemini Interactions API migrations, remove deprecated temperature, top_p, top_k, and candidate_count settings, replace thinking_budget with thinking_level, and remove prefilled model turns.

  • Keep multimodal expectations precise: images, audio, video, and PDFs are input formats, while the response is text.

  • Do not assume every upstream Google tool is automatically available through an OpenAI-compatible route. Check the live GPTProto API documentation for supported parameters, tools, file handling, and response fields.

  • Run a canary test with your own coding tasks, schemas, tool calls, latency targets, and token budgets before moving all requests.

This migration guidance is a major difference between a basic model listing and a usable Gemini 3.7 Flash API access page: it tells developers which existing requests can fail even when the new model ID is correct.

When Should You Choose Gemini 3.7 Flash?

Choose Gemini 3.7 Flash when a workload needs a long context window, multiple input formats, stronger first-pass coding, or multi-step tool orchestration at a lower token price than larger flagship models. Keep Gemini 3.6 Flash when minimal thinking is important or when an existing application has not yet passed its regression tests.

For Grok 4.6 vs Gemini 3.7 Flash, the decision is more specific. Gemini accepts audio, video, and PDF input and offers more than twice Grok's context window. Grok 4.6 adds an xhigh reasoning option and SpaceXAI documents no fixed text-output limit. Neither model is universally better, so route representative tasks through both before selecting a default.

Decision factor Gemini 3.7 Flash Gemini 3.6 Flash Grok 4.6
Best fit Coding, design adherence, multimodal documents, multi-step agents Existing Flash workflows and minimal-thinking traffic Coding and agents needing xhigh reasoning or very long text output
Context window 1,048,576 tokens 1,048,576 tokens 500,000 tokens
Native inputs Text, image, video, audio, PDF Text, image, video, audio, PDF Text, image
Text output limit 65,536 tokens 65,536 tokens No fixed limit documented by SpaceXAI
Reasoning control Low, medium, high Minimal, low, medium, high Low, medium, high, xhigh
Official short-context token price* $0.75 input / $3.75 output through Dec. 31, 2026 $0.75 input / $3.75 output through Dec. 31, 2026 $2 input / $6 output below 200K input tokens

*Official provider pricing is shown only for model selection context. Use GPTProto's live pricing panel for the rate billed through GPTProto.

Gemini 3.7 Flash API: Common Technical Questions

How much does the Gemini 3.7 Flash API cost on GPTProto?

GPTProto displays a 40% off rate in the live pricing panel. Use that panel for the current input and output prices because provider promotions and routing rates can change. Do not use Google's native API price as a substitute for the GPTProto billing rate.

How can I get a Gemini 3.7 Flash API key?

Create a GPTProto account, add balance, and generate one API key from the dashboard. The same key and balance can be used across Gemini 3.7 Flash and 200+ other supported models, so you do not need a separate credential for each provider.

What is the Gemini 3.7 Flash context window?

The model accepts up to 1,048,576 input tokens and can return up to 65,536 output tokens. A large advertised window does not guarantee that every retrieval or reasoning task will be accurate at the limit, so test with documents and repository structures that match your real workload.

Does Gemini 3.7 Flash support image, video, audio, and PDF input?

Yes. It accepts text, images, video, audio, and PDFs as input and returns text. It does not natively generate images or audio, and it does not support Google's Live API.

Is Gemini 3.7 Flash good for coding and AI agents?

Yes. Google positions it for coding and agents, and its launch results show gains over 3.6 Flash on FrontierCode, DeepSWE, WebDev Arena, and AutomationBench. Validate it on your own repositories, tool definitions, and acceptance tests before replacing a production model.

What changed from Gemini 3.6 Flash to Gemini 3.7 Flash?

The context and output limits remain the same, but 3.7 improves coding, design adherence, document reasoning, and multi-step execution in Google's evaluations. The main configuration difference is that 3.7 removes the minimal thinking level.

When was the Gemini 3.7 Flash API released?

Google released Gemini 3.7 Flash as a generally available model on August 13, 2026. The stable official model ID is gemini-3.7-flash, with no preview suffix.

Can I use Google's native Gemini SDK examples with a GPTProto key?

Not without changing the integration. Native Google SDK examples use Google's authentication, endpoint, and request schema. GPTProto uses its own documented, OpenAI-compatible access route, so follow the live GPTProto Quick Start and parameter reference.

Which is better: Grok 4.6 or Gemini 3.7 Flash?

Gemini 3.7 Flash is the clearer fit for 1M-context work and audio, video, or PDF understanding. Grok 4.6 is worth testing when xhigh reasoning or unrestricted text-output length matters. For coding quality, compare both on the same repository tasks rather than choosing from one benchmark or token price alone.

Related Articles

Guides, comparisons, and updates related to this model.

All Articles
What Is GLM-5.3? Z.ai's Quiet Coding Plan Launch, Pricing, and Confirmed Upgrades

What Is GLM-5.3? Z.ai's Quiet Coding Plan Launch, Pricing, and Confirmed Upgrades

GLM-5.3 is live in Z.ai’s Coding Plan. See its release status, 1M context, reasoning modes, pricing, upgrades, and remaining unknowns.

7 Best Image Editing AI Models in 2026 for API, Batch Editing, and Product Photos

7 Best Image Editing AI Models in 2026 for API, Batch Editing, and Product Photos

Compare 7 image editing AI models for APIs, batch workflows, product photos, text edits, and brand consistency, with the best pick for each task.

DeepSeek V4 Pro vs Kimi K3: What Changed After the 0813 Update?

DeepSeek V4 Pro vs Kimi K3: What Changed After the 0813 Update?

Compare DeepSeek V4 Pro 0813 vs Kimi K3 for coding, speed, multimodal input, and API cost—and see which model better fits your project.

Grok 4.6 vs DeepSeek V4 Pro: Coding, Pricing, and Which Is Better?

Grok 4.6 vs DeepSeek V4 Pro: Coding, Pricing, and Which Is Better?

Grok 4.6 vs DeepSeek V4 Pro compared on coding, frontend work, benchmarks, context and API pricing. See which model offers better value for developers.

GPT Proto

Empowering AI Innovation with Global Scale and Stability:

With our flagship product GPT Proto, we offer a unified interface to access and combine APIs from the world's leading AI providers—spanning text, vision, speech, and beyond. We empower developers and enterprises to simplify integration and accelerate innovation without limits.

Global Infrastructure, Local Compliance:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

Built to Scale:

We understand that stability is paramount. Our platform is built on a robust, decentralized architecture supporting dynamic Auto-scaling. Whether you are running a pilot or handling millions of concurrent requests, our system expands instantly to meet demand—guaranteeing that your business never outgrows our infrastructure.

Navigation

  • Dashboard
  • Models
  • Create Image
  • AI Image Upscale
  • AI Background Remover
  • Create Video
  • Edit in Canvas
  • Features
  • Pricing
  • AI Docs
  • AI Blog
  • AI Insight
  • AI Skills

Features

  • AI Packaging Design Generator
  • Anime to Real Life AI
  • Anime AI Art Generator
  • AI Object Remover
  • AI Image Editor
  • AI Motion Transfer
  • AI Watermark Remover
  • AI Image Enhancer Online
  • Online Background Remover Tool
  • AI Face Swap Image
  • AI Passport Photo Maker
  • MS Paint AI Generator
  • AI Clothes Remover
  • Unrestricted AI Image Generator
Explore all features >

LLM

  • Gemini 3.7 Flash
  • Grok 4.6
  • Qwen3.8 Max
  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
  • GPT 5.6 Sol
  • Grok 4.5
  • Claude Sonnet 5
  • Minimax M3
  • GLM 5.2
  • GPT 5.1 Chat Latest
  • Claude Fable 5
  • Qwen3.7 Max
  • Claude Opus 4.8 Thinking
  • Claude Opus 4.8
  • Gemini 3.5 Flash
Explore all models >

Image

  • Dola Seedream 5.0 Pro 260628
  • Gemini 3.1 Flash Lite Image
  • Gemini 3.1 Flash Image
  • GPT Image 2
  • Gemini 3.1 Flash Image Preview
  • Seedream 5.0 260128
  • Doubao Seedream 5.0 260128
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image O1
  • GPT Image 1.5
  • Seedream 4.5 251128
  • Doubao Seedream 4.5 251128
  • Grok Imagine 0.9
  • Gemini 3 Pro Image Preview
  • Qwen Image Lora
  • Qwen Image Plus Lora
  • Qwen Image Plus
  • Grok 4 Image
  • GPT Image 1 Mini
Explore all models >

Video

  • Dreamina Seedance 2.5 260628
  • Kling v3.0 4k
  • Dreamina Seedance 2.0 Mini 260615
  • Kling v3 Omni 4k
  • Dreamina Seedance 2.0 Fast 260128
  • Dreamina Seedance 2.0 260128
  • Vidu 2.0
  • Doubao Seedance 2.0 260128
  • Doubao Seedance 2.0 Fast 260128
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Vidu Q3 Turbo
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
  • Vidu Q2 Pro Fast
  • Vidu Q2
Explore all models >

© 2026 Talent Tech Global Limited (Hong Kong). All rights reserved.

Registered Address: Unit 1022a, Beverley Commercial Centre, 87-105 Chatham Road South, Tsim Sha Tsui, Hong KongCertificate No.: 79462435-000-12-25-0
  • About Us
  • Privacy Policy
  • Terms of Service
  • Sitemap