Gemini 3.5 Flash-Lite
Google's cheapest Gemini 3.5 tier, built for high-volume automation
Gemini 3.5 Flash-Lite shipped July 21, 2026 alongside Gemini 3.6 Flash as the cheapest model in the current Gemini lineup, aimed at teams running high-volume, latency-sensitive workloads like classification and bulk summarization rather than complex reasoning. It shares the 1,048,576 token context window with its larger Flash siblings and takes text, image, audio, and video input. Pricing runs $0.30 per million input tokens and $2.50 per million output tokens, with a flat 50 percent discount in batch mode for non-real-time jobs. Google positions it as the entry point into the Gemini 3.5 family for developers who need to process large volumes of requests cheaply and can route only the harder cases up to Flash or Pro.