OpenAI Expands API Access to GPT-4 Turbo with Vision

OpenAI makes its GPT-4 Turbo with Vision model generally available through its API, allowing developers to analyze images and text in a single call. The update streamlines app development for enterprises by integrating vision capabilities directly into JSON and function calling.

OpenAI releases GPT-4 Turbo with Vision for general availability through its application programming interface, giving enterprise developers broader access to its powerful multimodal capabilities. This update combines the speed, larger context windows, and affordability of the GPT-4 Turbo model with the ability to process and understand images.

Developers now process text and images through a single API call instead of relying on separate models, which significantly streamlines application workflows. Furthermore, the vision capabilities integrate directly into JSON requests and function calling, allowing connected applications to automatically trigger actions like sending emails or making purchases based on image analysis.

Several startups already utilize this technology to power innovative features in their products. Cognition's autonomous coding agent Devin uses the model to generate code, Healthify relies on it to analyze meal photos for nutritional insights, and TLDraw transforms user drawings into functional websites.

Read More at the original source →