OpenAI Releases Upgraded GPT-4-Turbo With Vision Capabilities for Developers

OpenAI introduces a significantly improved GPT-4-Turbo model that features advanced vision technology for analyzing images and video. The update is currently available to developers through the API and will soon arrive in the consumer ChatGPT application.

OpenAI releases a majorly improved version of its GPT-4-Turbo artificial intelligence model, offering enhanced response quality and advanced analysis capabilities. This updated model prominently features AI vision technology that allows the system to understand and process content from images, video, and audio inputs. For the first time, third-party developers gain access to GPT-4-Turbo with vision through the API, paving the way for innovative new applications in fashion, coding, and gaming.

The latest update extends the AI's knowledge cutoff date to December 2023, replacing the previous April 2023 limit so the model provides more current information. OpenAI emphasizes that this release significantly streamlines developer workflows by combining image and text processing into a single model rather than requiring separate systems. Additionally, developers can now utilize JSON mode and function calling alongside these new vision requests to build more efficient software.

While the upgraded GPT-4-Turbo remains exclusive to the developer API for now, OpenAI confirms that these powerful vision capabilities will soon arrive in the consumer ChatGPT application. This planned expansion positions OpenAI to better compete with rivals like Google, which recently integrated similar multimodal features into its Gemini Pro 1.5 platform. As the technology rolls out to everyday users, people will experience much more efficient image and video understanding directly inside ChatGPT.

Read More at the original source →