Google Makes Gemini 2.5 Flash and Pro Generally Available, Launches Flash-Lite

Google brings Gemini 2.5 Flash and Pro models to general availability on Vertex AI while introducing a cost-effective Flash-Lite preview. The updates give enterprise developers production-ready tools for advanced reasoning and high-volume AI tasks.

Google expands the Gemini 2.5 model family on Vertex AI by making both Gemini 2.5 Flash and Gemini 2.5 Pro generally available. These production-ready models provide enterprise developers with the stability and scalability needed to deploy advanced AI capabilities into mission-critical applications. The general availability launch allows organizations to confidently use these high-speed models for complex reasoning tasks at scale.

In addition to the stable releases, Google introduces Gemini 2.5 Flash-Lite in public preview as the most cost-efficient model in the 2.5 series. This new lightweight model targets high-volume, cost-sensitive operations such as classification, translation, and intelligent routing. Furthermore, Google makes Supervised Fine-Tuning generally available for Gemini 2.5 Flash, enabling businesses to tailor the model directly to their unique enterprise data and specific operational needs.

Google also updates the Live API with native audio capabilities in public preview to streamline the development of real-time audio AI systems. Together, these updates give enterprise builders a comprehensive toolkit to create sophisticated, customized, and efficient AI solutions. From large-scale summarization to responsive chat applications, these new milestones continue the rapid momentum of the Gemini 2.5 era across Google's AI platforms.

Read More at the original source →