ChatGPT Gains Voice and Image Capabilities in Major Update
OpenAI expands ChatGPT beyond text by adding voice conversation and image analysis features for paid users. The update draws mixed reactions as it pushes the AI closer to human-like assistants.
OpenAI brings voice and image capabilities to ChatGPT, expanding the platform beyond its traditional text-based prompts. The company rolls out these new features over two weeks to paid subscribers, with free users gaining access shortly after. This update transforms the chatbot into a more comprehensive assistant that competes directly with popular voice tools like Apple's Siri and Amazon's Alexa.
With the new voice feature, users engage in spoken conversations with the AI, allowing it to narrate bedtime stories, settle dinner table debates, and read text aloud. Spotify already utilizes this underlying technology to help podcasters translate their content into different languages. Additionally, the image feature enables users to upload photos, highlight specific areas with a drawing tool, and ask the AI to troubleshoot broken appliances, plan meals based on fridge contents, or analyze complex work graphs.
The announcement sparks mixed reactions online, blending excitement with skepticism. UC Berkeley professor Trevor Darrell warns that making the chatbot sound more human risks falling into the "uncanny valley gap," where interfaces that almost mimic human interaction feel strange and off-putting. Despite these concerns about usability, the update marks a significant leap in how users interact with AI technology on a daily basis.