Sub-100ms Latency
Optimized for instant-feel applications with a 40% faster TTFT than standard Gemini Flash models.

image
text
Technical highlights of the Gemini 3.5 Flash Lite architecture.
Optimized for instant-feel applications with a 40% faster TTFT than standard Gemini Flash models.

Analyze images, audio, and video natively with 68.2% MMMU accuracy for complex vision-based tasks.

The most affordable Gemini model at $0.075/1M tokens, perfect for high-throughput data labeling.

Process massive datasets or hour-long videos with the signature Gemini 1,048,576 token capacity.

Getting a gemini-3.5-flash-lite API key takes four steps and a few minutes. Create a free GPTProto account, add credits, generate your key, and make your first call — at $0.18 / $1.5 it's a cheaper gemini-3.5-flash-lite API key than going direct, and one key works across every model on the platform. Full gemini-3.5-flash-lite Documentation is in the docs.

Sign up

Top up

Generate your API key

Make your first API call
Answers to common questions about the Gemini 3.5 Flash Lite model features, pricing, and technical performance.