Google Unveils Gemini, a New Multimodal AI Model Family
Google releases Gemini, its most advanced AI model capable of processing text, images, audio, and video simultaneously. The technology launches in three optimized sizes to handle tasks ranging from on-device operations to highly complex reasoning.
Google unveils Gemini, introducing its most capable artificial intelligence model to date. Developed by Google DeepMind, this new technology represents a significant leap forward in the AI industry by seamlessly processing and combining text, code, audio, image, and video inputs. CEO Sundar Pichai highlights this release as a profound technological shift that surpasses the historical impact of the mobile and web revolutions.
The Gemini 1.0 release features three distinct model variants tailored for specific applications. Gemini Ultra handles highly complex tasks and sets new state-of-the-art standards across 30 of 32 industry benchmarks. Gemini Pro serves as a versatile option for scaling across a wide range of everyday tasks, while Gemini Nano provides efficient processing for on-device applications.
Google immediately integrates this new technology into its products by updating Bard with the Gemini Pro model. This upgrade brings advanced reasoning, expert coding skills, and enhanced multimodal understanding directly to users. The launch solidifies Google's ongoing commitment to its AI-first vision and drives new levels of innovation, creativity, and productivity.