OpenAI Unveils GPT-4o, Bringing Free Voice and Vision to ChatGPT

OpenAI introduces GPT-4o, a powerful new AI model that turns ChatGPT into a real-time, multimodal voice assistant available to all users. The release strategically precedes major AI announcements from rivals Google and Apple.

OpenAI unveils GPT-4o, a major upgrade to its artificial intelligence that makes ChatGPT smarter and easier to use. Unlike the previous GPT-4 model, this new version is available to unpaid customers, giving everyone free access to the company's most advanced technology. Chief Technology Officer Mira Murati introduces the update as a massive leap forward in user experience.

The new model effectively transforms ChatGPT into a digital personal assistant capable of engaging in real-time, spoken conversations. GPT-4o is multimodal, meaning it seamlessly processes text, audio, and vision to analyze uploaded photos, documents, or charts. Additionally, the system features memory capabilities to learn from past interactions and performs real-time language translation.

This highly anticipated release arrives as OpenAI attempts to maintain its lead in the fiercely competitive AI arms race. The timing strategically precedes Google's annual I/O developer conference and expected AI announcements from Apple next month. Furthermore, this latest advancement provides a significant advantage to Microsoft, which relies heavily on OpenAI technology to power its own software products.

Read More at the original source →