Predictive APIs and Machine Learning Integration
Gemini 2.5 Flash-Lite is now stable and generally available- Google Developers Blog

Gemini 2.5 Flash-Lite is now stable and generally available- Google Developers Blog

7/22/2025 · Logan Kilpatrick, Zach Gleicher

What this post added

This post announces the stable and general availability of Gemini 2.5 Flash-Lite, a new, cost-efficient, and fast model within the Gemini 2.5 family. It highlights its lower latency compared to previous versions, its pricing ($0.10 input per 1M, $0.40 output per 1M tokens), and its improved quality across benchmarks. Key features include a 1 million-token context window, controllable thinking budgets, and native tool support (Grounding with Google Search, Code Execution, URL Context). The post also provides examples of its successful deployment in various applications, demonstrating its impact on latency reduction, power consumption, and content processing.

Read the original post ↗