OpenAI Launches GPT-4o Omnimodel for Real-Time Voice and Video
OpenAI introduces GPT-4o, a unified AI model that enables real-time voice, video, and text interactions for free. The new omnimodel delivers faster, more natural conversations that rival traditional virtual assistants.
OpenAI unveils GPT-4o, a new flagship AI model that allows users to interact in real time through live voice, video, and text. The company rolls out this "omnimodel" for free to all users via its app and web interface, while paid subscribers receive higher request limits.
Unlike previous versions that relied on separate models for different inputs, GPT-4o merges text, audio, and visual capabilities into a single system. This unified approach significantly reduces response times and enables much smoother transitions between complex tasks.
Demonstrations highlight the model's ability to handle natural, fluid conversations where users interrupt, change the AI's tone, or request dramatic readings. The result functions like a highly advanced version of Siri or Alexa, marking a major shift toward more natural human-machine collaboration.