GPT Proto

GPTProto

  • Dashboard
  • LLM

    • openai
      GPT 6 AstraNew
    • z-ai
      GLM 5.3
    • claude
      Claude Fable 5
    • deepseek
      DeepSeek v4 Pro
    • google
      Gemini 3.7 Flash
    • grok
      Grok 4.6
    Explore models >

    Image

    • bytedance
      Seedream 5.0 Pro (Build 260628)New
    • openai
      GPT Image 2
    • google
      Nano Banana Pro (Gemini 3 Pro Image)
    • google
      Nano Banana 2 (Gemini 3.1 Flash Image)
    • midjourney
      Midjourney
    • google
      Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image)
    Explore models >

    Video

    • qwen
      Wan 3.0New
    • bytedance
      Seedance 2.5 (Build 260628)
    • bytedance
      Seedance 2.0 (Build 260128)
    • bytedance
      Seedance 2.0 Mini (Build 260615)
    • kling
      Kling v3.0 4K
    • vidu
      Vidu Q3 Turbo
    Explore models >
    Explore 226+ Models >
  • Generator

    • Create Image
    • Create Video
    • Edit in Canvas
    • Chat

    Features

    • Cute Wallpaper GeneratorNew
    • AI French Kissing Generator
    • AI Age Filter
    • AI Packaging Design Generator
    • Anime to Real Life AI
    • Anime AI Art Generator
    • AI Object Remover
    • AI Image Editor
    • AI Motion Transfer
    • AI Watermark Remover
    Explore All >

    Prompts

    • Seedance 2.0 PromptsNew
    • GPT Image 2 Prompts
    • Nano Banana Pro Prompts
    • Seedream 5.0 Pro Prompts
    • Midjourney Prompts
  • AI Blog

    • 6 Cheapest AI Image Generators in 2026: Real Cost per Image
    • Claude vs ChatGPT for Coding in 2026: Which Is Better for Debugging, Frontend, Python, and Large Codebases?
    • GLM 5.3 Flash vs DeepSeek V4 Flash: Which Is Better for Code, Agents, and Cost?
    • 5 Best Midjourney API Alternatives in 2026: Real Model APIs, Not Discord Wrappers
    • Qwen3.8-Flash-Next vs GLM-5.3 Flash: Which Is Better for Coding, Agents, and Price?
    Explore All >

    AI Insight

    • Fable 5.1 vs Opus 5: Best AI for Agentic Coding
    • Introducing Claude Fable 5.1 and Claude Mythos 5.1: Same Model, Different Safeguards
    • What Is Hunyuan 4? Tencent Hy4 Preview Features, Pricing, Benchmarks, and Release Status
    • Can Nano Banana Generate Multiple Images at Once?
    • how to reduce the claude token usage effectively
    Explore All >

    AI Docs

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    Explore All >

    AI Skills

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    Explore All >
Pricing+7% bonus
English繁體中文한국어日本語EspañolРусский
Get Started Now
  1. Home
  2. /Model
  3. /Google
  4. /gemini-3.5-flash-lite / web-search
Google
Gemini 3.5 Flash Lite
$ 
Gemini 3.5 Flash Lite API offers sub-second latency and a massive 1M token context window. Optimized for high-frequency tasks, this Google model provides multimodal support and elite JSON extraction at a fraction of the cost of standard models.

Modalities

Input: TextInput: ImageInput: Document
Output: Text

/

Context

API Usage Examples
$ 
curl --request POST "https://gptproto.com/v1/chat/completions" \
  --header "Authorization: Bearer $GPTPROTO_API_KEY" \
  --header "Content-Type: application/json" \
  --data '{
    "model": "gemini-3.5-flash-lite",
    "messages": [
      {
        "role": "user",
        "content": "Hello"
      }
    ]
  }'
Gemini 3.5 Flash Lite pricing

Estimate a request with real work scenarios. GPTProto token pricing is 40% below official rates.

UsageQuantityRateCost
tokens
$0.18/1M$0.0002
tokens
$1.5/1M$0.0011
tokens
$0.6/1M$0.0018
tokens
$0.018/1M$0.0004
Cost per request$0.0037
Requests
Top-up amount

Top-up $100 and you get:

1.

Top-up credits with permanent validity. You will receive a total of $100.00.

2.

Additional 40% model discount, saving $66.6621 versus direct official Google API calls.

Related Models
All Models
Gemini 3.5 Flash Lite
Current
$ 
byGoogle1.05M context$0.18/M input$1.5/M output
GPT 6 Astra
$ 
byOpenAI$8/M input$40/M output
Gemini 3.8 Flash
$ 
byGoogle1.05M context$0.9/M input$4.5/M output
Claude Fable 5.1
$ 
byClaude1M context$9/M input$45/M output
Qwen3.8 Max 0902
$ 
byQwen$1.8/M input$5.4/M output
GLM 5.3 Flash
$ 
byZ-AI1.31M context$0.15/M input$0.5/M output
DeepSeek v4 Flash Vision Exp
$ 
byDeepSeek1.05M context$0.44/M input$1.32/M output
GLM 5.3
$ 
byZ-AI1.31M context$1.26/M input$3.96/M output
Gemini 3.7 Flash
$ 
byGoogle1.05M context$0.9/M input$4.5/M output
Grok 4.6
$ 
byGrok500K context$1.2/M input$3.6/M output
Qwen3.8 Max
$ 
byQwen1M context$1.8/M input$5.4/M output
Claude Opus 5
$ 
byClaude1M context$4.5/M input$22.5/M output
Gemini 3.6 Flash
$ 
byGoogle1.05M context$0.9/M input$4.5/M output
Kimi K3
$ 
byMoonshotAI1.05M context$2.7/M input$13.5/M output
GPT 5.6 Luna
$ 
byOpenAI1.05M context$0.16/M input$0.96/M output
GPT 5.6 Terra
$ 
byOpenAI1.05M context$1.6/M input$9.6/M output
Grok 4.5
$ 
byGrok500K context$1.2/M input$3.6/M output
Claude Sonnet 5
$ 
byClaude1M context$1.8/M input$9/M output
MiniMax M3
$ 
byMiniMax1.05M context$0.48/M input$0.96/M output
GLM 5.2
$ 
byZ-AI1.05M context$1.26/M input$3.96/M output
Qwen3.7 Max
$ 
byQwen1M context$0.36/M input$1.44/M output
Gemini 3.5 Flash
$ 
byGoogle1.05M context$0.9/M input$5.4/M output
DeepSeek v4 Flash
$ 
byDeepSeek1.05M context$0.44/M input$1.32/M output
DeepSeek v4 Pro
$ 
byDeepSeek1.05M context$1.32/M input$3.96/M output
Grok 4.3
$ 
byGrok1M context$0.75/M input$1.5/M output
Kimi K2.6
$ 
byMoonshotAI262K context$0.855/M input$3.6/M output
Gemini 3.1 Flash Lite Preview
$ 
byGoogle1.05M context$0.15/M input$0.9/M output
MiniMax M2.5
$ 
byMiniMax205K context$0.24/M input$0.96/M output
Gemini 3.1 Pro Preview
$ 
byGoogle1.05M context$1.2/M input$7.2/M output
Kimi K2.5
$ 
byMoonshotAI262K context$0.54/M input$2.7/M output
Doubao Seed 1.6 Thinking (Build 250715)
$ 
byBytedance262K context$0.0971/M input$0.9714/M output
Doubao Seed 1.6 Thinking (Build 250615)
$ 
byBytedance262K context$0.0971/M input$0.9714/M output
Doubao Seed 1.6 Flash (Build 250615)
$ 
byBytedance262K context$0.0182/M input$0.1821/M output
ModelInput → Output
Gemini 3.5 Flash LiteCurrent
$ 
1.05M$0.18 / $1.50 per 1M$0.60 / $0.02 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GPT 6 Astra
$ 
—$8.00 / $40.00 per 1M$10.00 / $0.80 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Gemini 3.8 Flash
$ 
1.05M$0.90 / $4.50 per 1M$0.60 / $0.09 per 1M
Input: TextInput: ImageInput: VideoInput: DocumentInput: Audio
Output: Text
Claude Fable 5.1
$ 
1M$9.00 / $45.00 per 1M$11.25 / $0.23 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Qwen3.8 Max 0902
$ 
—$1.80 / $5.40 per 1M$2.25 / $0.23 per 1M
Input: TextInput: ImageInput: VideoInput: DocumentInput: Audio
Output: Text
GLM 5.3 Flash
$ 
—1.31M$0.15 / $0.50 per 1M— / $0.03 per 1M
Input: TextInput: ImageInput: VideoInput: Document
Output: Text
DeepSeek v4 Flash Vision Exp
$ 
—1.05M$0.44 / $1.32 per 1M— / $0.01 per 1M
Input: TextInput: Image
Output: Text
GLM 5.3
$ 
1.31M$1.26 / $3.96 per 1M— / $0.23 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Gemini 3.7 Flash
$ 
1.05M$0.90 / $4.50 per 1M$0.60 / $0.09 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Grok 4.6
$ 
500K$1.20 / $3.60 per 1M— / $0.30 per 1M
Input: TextInput: Image
Output: Text
Qwen3.8 Max
$ 
1M$1.80 / $5.40 per 1M$2.25 / $0.23 per 1M
Input: TextInput: ImageInput: VideoInput: Document
Output: Text
Claude Opus 5
$ 
1M$4.50 / $22.50 per 1M$5.63 / $0.45 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Gemini 3.6 Flash
$ 
1.05M$0.90 / $4.50 per 1M$0.60 / $0.09 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Kimi K3
$ 
1.05M$2.70 / $13.50 per 1M$0.27 / $0.27 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GPT 5.6 Luna
$ 
1.05M$0.16 / $0.96 per 1M$0.20 / $0.02 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GPT 5.6 Terra
$ 
1.05M$1.60 / $9.60 per 1M$2.00 / $0.16 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Grok 4.5
$ 
500K$1.20 / $3.60 per 1M$0.30 / $0.30 per 1M
Input: TextInput: Image
Output: Text
Claude Sonnet 5
$ 
1M$1.80 / $9.00 per 1M$2.25 / $0.18 per 1M
Input: TextInput: Document
Output: Text
MiniMax M3
$ 
1.05M$0.48 / $0.96 per 1M$0.10 / $0.10 per 1M
Input: TextInput: ImageInput: Document
Output: Text
GLM 5.2
$ 
1.05M$1.26 / $3.96 per 1M$0.23 / $0.23 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Qwen3.7 Max
$ 
1M$0.36 / $1.44 per 1M$0.07 / $0.07 per 1M
Input: TextInput: Document
Output: Text
Gemini 3.5 Flash
$ 
1.05M$0.90 / $5.40 per 1M$0.60 / $0.09 per 1M
Input: TextInput: ImageInput: Document
Output: Text
DeepSeek v4 Flash
$ 
—1.05M$0.44 / $1.32 per 1M— / $0.01 per 1M
Input: Text
Output: Text
DeepSeek v4 Pro
$ 
—1.05M$1.32 / $3.96 per 1M— / $0.04 per 1M
Input: Text
Output: Text
Grok 4.3
$ 
1M$0.75 / $1.50 per 1M$0.12 / $0.12 per 1M
Input: TextInput: Image
Output: Text
Kimi K2.6
$ 
262K$0.85 / $3.60 per 1M$0.14 / $0.14 per 1M
Input: TextInput: Document
Output: Text
Gemini 3.1 Flash Lite Preview
$ 
1.05M$0.15 / $0.90 per 1M$0.60 / $0.01 per 1M
Input: TextInput: ImageInput: Document
Output: Text
MiniMax M2.5
$ 
205K$0.24 / $0.96 per 1M$0.30 / $0.02 per 1M
Input: TextInput: Document
Output: Text
Gemini 3.1 Pro Preview
$ 
1.05M$1.20 / $7.20 per 1M$2.70 / $0.12 per 1M
Input: TextInput: ImageInput: Document
Output: Text
Kimi K2.5
$ 
262K$0.54 / $2.70 per 1M$0.09 / $0.09 per 1M
Input: TextInput: Document
Output: Text
Doubao Seed 1.6 Thinking (Build 250715)
$ 
262K$0.10 / $0.97 per 1M—
Input: TextInput: Image
Output: Text
Doubao Seed 1.6 Thinking (Build 250615)
$ 
262K$0.10 / $0.97 per 1M—
Input: TextInput: Image
Output: Text
Doubao Seed 1.6 Flash (Build 250615)
$ 
262K$0.02 / $0.18 per 1M—
Input: TextInput: Image
Output: Text

Key Gemini 3.5 Flash Lite API Capabilities

Explore the core technical strengths of the Gemini 3.5 Flash Lite API, from its massive memory to its industry-leading speed, designed specifically for high-efficiency enterprise deployments.

Sub-100ms Response Speed

Optimized for an instant-feel, the model delivers a Time To First Token approximately 30-40% faster than standard versions, ensuring smooth user experiences in chat and real-time tools.

1M Token Massive Context

Ingest entire technical manuals or legal archives at once. The signature 1M context window allows for large-scale RAG and deep document analysis at a fraction of the price of larger models.

Native Multimodal Support

Process images, video up to one hour, audio, and PDFs natively. Gemini 3.5 Flash-Lite achieves a 68.2% score on MMMU benchmarks, outperforming many text-only lite competitors significantly.

Structured JSON Reliability

Features an optimized internal head for schema adherence, resulting in a 98% success rate on complex nested extraction. Perfect for converting unstructured data into actionable insights.

Gemini 3.5 Flash Lite API: Common Questions

Get details on the Gemini 3.5 Flash Lite API. Learn about latency, the 1M context window, pricing, and how this model compares to others in the 3.5 family for building responsive AI applications.

What is the typical latency of the Flash Lite model?

Gemini 3.5 Flash-Lite is optimized for instant-feel applications. It delivers a Time To First Token (TTFT) that is roughly 30-40% faster than the standard Flash model on text-only prompts. For most short queries, you can expect a response within 100-200ms, making it ideal for interactive chatbots and real-time translation where latency is the most critical factor for ensuring a smooth and responsive user experience.

How large is the Gemini 3.5 Flash Lite context window?

This model features a massive 1,048,576 token context window. This allows you to process enormous amounts of data—such as hour-long videos, large codebases, or hundreds of documents—in a single request. This high efficiency enables large-scale Retrieval-Augmented Generation (RAG) at a significantly lower price point than larger models, without needing to constantly truncate or summarize your input data before processing.

What is the pricing for the Gemini 3.5 Flash Lite API?

The pricing is highly competitive at $0.075 per 1M input tokens and $0.30 per 1M output tokens. It is approximately 50% cheaper than the standard Gemini 3.5 Flash. Additionally, we offer a 50% discount on cached input tokens. This makes it the most cost-effective model per token in the entire Gemini 3.5 family, perfect for high-volume data labeling, classification, or massive document summarization pipelines.

Does this lite model support multimodal inputs?

Yes. Unlike many competitors that offer text-only lite versions, this model natively supports multimodal inputs including images, video (up to 1 hour), audio, and PDF files. It achieves an impressive 68.2% on the MMMU benchmark. This capability allows you to build sophisticated applications that can search video archives, index image libraries, or analyze complex technical documents containing charts and graphs effortlessly.

How does it compare to GPT-4o-mini?

While GPT-4o-mini is strong in reasoning, Gemini 3.5 Flash-Lite offers a much larger context window (1M vs 128k) and lower latency (120ms vs 180ms). It is also roughly 50% cheaper on input tokens. For tasks like large-scale document summarization, high-volume classification, or multimodal processing, Gemini 3.5 Flash-Lite provides a superior balance of speed, capacity, and cost-efficiency for modern developers.

Is my data used to train the Gemini models?

No. We take data privacy seriously. Any data you send via the GPTProto.com platform to the Gemini 3.5 Flash-Lite API is not used to train the underlying foundation models. We provide a unified OpenAI-compatible endpoint with consolidated billing and automatic failover across multiple Google Cloud regions to ensure higher uptime, allowing you to build enterprise-grade applications with full confidence in your data security.

GPT Proto

Empowering AI Innovation with Global Scale and Stability:

With our flagship product GPT Proto, we offer a unified interface to access and combine APIs from the world's leading AI providers—spanning text, vision, speech, and beyond. We empower developers and enterprises to simplify integration and accelerate innovation without limits.

Global Infrastructure, Local Compliance:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

Built to Scale:

We understand that stability is paramount. Our platform is built on a robust, decentralized architecture supporting dynamic Auto-scaling. Whether you are running a pilot or handling millions of concurrent requests, our system expands instantly to meet demand—guaranteeing that your business never outgrows our infrastructure.

Navigation

  • Dashboard
  • Models
  • Create Image
  • AI Image Upscale
  • AI Background Remover
  • Create Video
  • Edit in Canvas
  • Chat
  • Features
  • Pricing
  • AI Docs
  • AI Blog
  • AI Insight
  • AI Skills

Features

  • Cute Wallpaper Generator
  • AI French Kissing Generator
  • AI Age Filter
  • AI Packaging Design Generator
  • Anime to Real Life AI
  • Anime AI Art Generator
  • AI Object Remover
  • AI Image Editor
  • AI Motion Transfer
  • AI Watermark Remover
  • AI Image Enhancer Online
  • Online Background Remover Tool
  • AI Face Swap Image
  • AI Passport Photo Maker
  • MS Paint AI Generator
  • AI Clothes Remover
  • Unrestricted AI Image Generator
  • AI French Kissing Generator
  • AI Movie Poster Generator
  • Artlist IO studio
Explore all features >

LLM

  • GPT 6 Astra
  • GLM 5.3
  • Claude Fable 5
  • DeepSeek v4 Pro
  • Gemini 3.7 Flash
  • Grok 4.6
  • Gemini 3.8 Flash
  • Claude Fable 5.1
  • Qwen3.8 Max 0902
  • GLM 5.3 Flash
  • DeepSeek v4 Flash Vision Exp
  • Qwen3.8 Max
  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
  • GPT 5.6 Sol
  • Grok 4.5
Explore all models >

Image

  • Seedream 5.0 Pro (Build 260628)
  • GPT Image 2
  • Nano Banana Pro (Gemini 3 Pro Image)
  • Nano Banana 2 (Gemini 3.1 Flash Image)
  • Midjourney
  • Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image)
  • Nano Banana 2 (Gemini 3.1 Flash Image)
  • Seedream 5.0 (Build 260128)
  • Doubao Seedream 5.0 (Build 260128)
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image o1
  • GPT Image 1.5
  • Seedream 4.5 (Build 251128)
  • Doubao Seedream 4.5 (Build 251128)
  • Grok Imagine 0.9
  • Qwen Image LoRA
  • Qwen Image Plus LoRA
  • Qwen Image Plus
  • Grok 4 Image
Explore all models >

Video

  • Wan 3.0
  • Seedance 2.5 (Build 260628)
  • Seedance 2.0 (Build 260128)
  • Seedance 2.0 Mini (Build 260615)
  • Kling v3.0 4K
  • Vidu Q3 Turbo
  • Kling v3 Omni 4K
  • Seedance 2.0 Fast (Build 260128)
  • Vidu 2.0
  • Doubao Seedance 2.0 (Build 260128)
  • Doubao Seedance 2.0 Fast (Build 260128)
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
  • Vidu Q2 Pro Fast
Explore all models >

Contact us

Questions or feedback? Reach us through any of the channels below.

TelegramWhatsApp

© 2026 Talent Tech Global Limited (Hong Kong). All rights reserved.

Registered Address: Unit 1022a, Beverley Commercial Centre, 87-105 Chatham Road South, Tsim Sha Tsui, Hong KongCertificate No.: 79462435-000-12-25-0
  • About Us
  • Privacy Policy
  • Terms of Service
  • Sitemap
Friendslogoto.videotopostudio.cc