Gemini 2.5 Flash-Lite is now stable and generally available
Blog post from Google Cloud
Gemini 2.5 Flash-Lite, the latest release in the Gemini 2.5 model family, combines high speed and cost-efficiency, priced at $0.10 per million input tokens and $0.40 per million output tokens, making it ideal for high-volume tasks like translation and classification. It offers a significant reduction in latency and power consumption compared to previous models, with improved performance across various benchmarks such as coding, math, and multimodal understanding. Users benefit from a 1 million-token context window and support for native tools like Google Search and Code Execution. Notable applications include Satlyt’s space computing platform, HeyGen’s multilingual video content automation, DocsHound’s documentation processing, and Evertune’s brand representation analysis. The model is now available for deployment in Google AI Studio and Vertex AI, with a transition from its preview version scheduled for August 25th.
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.