OpenAI Unveils Omni-Modal GPT-4o to Fend Off Rising AI Competitors
OpenAI introduces GPT-4o, a faster and cheaper AI model that handles text, audio, and video in real-time. The update brings premium features to free users just as Google prepares to host its own developer conference.
OpenAI introduces GPT-4o, a faster and cheaper upgrade to the AI model that powers ChatGPT, as the company strives to maintain its lead in a fiercely competitive market. The new "omni" model combines voice, text, and vision capabilities into a single system, enabling it to process and respond to audio inputs in milliseconds and generate images from visual prompts.
This update significantly levels the playing field by granting free users access to premium features that previously required a paid subscription. Free users now have the ability to search the web, converse with the chatbot using various voices, and instruct the AI to remember specific details for future interactions.
The launch intentionally precedes Google's I/O developer conference, signaling a bold move to overshadow upcoming announcements from the search giant and other rivals like Anthropic. Despite a highly publicized demo that features minor audio glitches and an unexpectedly flirtatious comment from the AI, OpenAI is currently rolling out the new text and image capabilities to paying users.