Google releases Gemini 2.5 Flash-Lite, its fastest and lowest cost AI model

Published 22/07/2025, 18:14
© Reuters.

Investing.com -- Google (NASDAQ:GOOGL) has released the stable version of Gemini 2.5 Flash-Lite, completing its lineup of 2.5 models ready for production use. The new model is designed to balance performance and cost while maintaining quality for latency-sensitive tasks like translation and classification.

Gemini 2.5 Flash-Lite offers lower latency than previous 2.0 Flash-Lite and 2.0 Flash models across a broad range of prompts. It is priced at $0.10 per million input tokens and $0.40 per million output tokens, making it Google’s lowest-cost 2.5 model. The company has also reduced audio input pricing by 40% from the preview launch.

Despite its cost efficiency, the model demonstrates higher quality than 2.0 Flash-Lite across benchmarks including coding, math, science, reasoning, and multimodal understanding. Users get access to a 1 million-token context window, controllable thinking budgets, and support for native tools like Grounding with Google Search, Code Execution, and URL Context.

Several companies have already implemented the new model. Satlyt, a decentralized space computing platform, has seen a 45% reduction in latency for critical onboard diagnostics and a 30% decrease in power consumption. HeyGen uses the model to automate video planning and translate videos into over 180 languages. DocsHound leverages it to process long videos and extract screenshots with low latency, while Evertune uses it to analyze how brands are represented across AI models.

Developers can start using Gemini 2.5 Flash-Lite by specifying "gemini-2.5-flash-lite" in their code. Google plans to remove the preview alias on August 25, 2025. The model is available in Google AI Studio and Vertex (NASDAQ:VRTX) AI.

This article was generated with the support of AI and reviewed by an editor. For more information see our T&C.

Latest comments

Risk Disclosure: Trading in financial instruments and/or cryptocurrencies involves high risks including the risk of losing some, or all, of your investment amount, and may not be suitable for all investors. Prices of cryptocurrencies are extremely volatile and may be affected by external factors such as financial, regulatory or political events. Trading on margin increases the financial risks.
Before deciding to trade in financial instrument or cryptocurrencies you should be fully informed of the risks and costs associated with trading the financial markets, carefully consider your investment objectives, level of experience, and risk appetite, and seek professional advice where needed.
Fusion Media would like to remind you that the data contained in this website is not necessarily real-time nor accurate. The data and prices on the website are not necessarily provided by any market or exchange, but may be provided by market makers, and so prices may not be accurate and may differ from the actual price at any given market, meaning prices are indicative and not appropriate for trading purposes. Fusion Media and any provider of the data contained in this website will not accept liability for any loss or damage as a result of your trading, or your reliance on the information contained within this website.
It is prohibited to use, store, reproduce, display, modify, transmit or distribute the data contained in this website without the explicit prior written permission of Fusion Media and/or the data provider. All intellectual property rights are reserved by the providers and/or the exchange providing the data contained in this website.
Fusion Media may be compensated by the advertisers that appear on the website, based on your interaction with the advertisements or advertisers
© 2007-2025 - Fusion Media Limited. All Rights Reserved.