Gemini 3.5 Flash

Google's default Flash model, running about four times faster than rival frontier models

Paid Multimodal
Visit Product Page →
1048576 context tokens
Undisclosed parameters
Proprietary license
May 2026 released

Gemini 3.5 Flash launched at Google I/O on May 20, 2026 and immediately replaced Gemini 3 Flash as the default model in the Gemini app, Google Search’s AI Mode, and the Gemini API. It handles a 1,048,576 token input window with up to 65,536 tokens of output across text, image, audio, and video, and Google says it runs about four times faster than other frontier models while beating Gemini 3.1 Pro on several coding and agentic benchmarks, including 76.2 percent on Terminal-Bench 2.1. Pricing is $1.50 per million input tokens and $9.00 per million output tokens, with cached input tokens discounted 90 percent to $0.15 per million, and dynamic thinking is enabled by default rather than requiring a separate flag.