On July 21, Google made a surprising move by unveiling two new models—Gemini 3.6 Flash and Gemini 3.5 Flash-Lite—which are now accessible via the Google AI Studio and Gemini platforms. Serving as the primary model, Gemini 3.6 Flash has demonstrated marked improvements in coding, reasoning, and tool invocation capabilities. It also reduces output token consumption by 17% to 65%, translating to lower operational costs. On the other hand, Gemini 3.5 Flash-Lite is designed for cost-efficiency and high-speed performance, boasting an output speed of up to 350 tokens per second, making it an ideal choice for high-throughput applications. In addition, Google has also rolled out the Gemini 3.5 Flash Cyber model, specifically tailored for cybersecurity purposes. Nevertheless, the launch date for the much-anticipated flagship model, Gemini 3.5 Pro, is still pending confirmation.
