OpenAI Launches Cost-Effective GPT-4o Mini to Replace GPT-3.5 Turbo

OpenAI releases GPT-4o mini, a faster and cheaper small AI model that outperforms competitors on reasoning benchmarks. The new model immediately replaces GPT-3.5 Turbo for developers and ChatGPT users.

OpenAI introduces GPT-4o mini, a new small AI model that is faster and more affordable than its larger counterparts. The model is available immediately for developers through the API and for consumers via the ChatGPT web and mobile apps, with enterprise access following next week. This release directly replaces GPT-3.5 Turbo as the smallest model in OpenAI's lineup.

According to OpenAI, GPT-4o mini outperforms competing small AI models like Gemini 1.5 Flash and Claude 3 Haiku on key reasoning and math benchmarks. The model scores 82% on the MMLU reasoning benchmark and 87% on the MGSM math reasoning test. It currently supports text and vision inputs through the API, with plans to add video and audio capabilities in the future.

The new model costs 15 cents per million input tokens and 60 cents per million output tokens, making it more than 60% cheaper than GPT-3.5 Turbo. It features a 128,000-token context window and targets high-volume, simple tasks where speed and cost efficiency are critical for developers. OpenAI frames this release as a major step toward making AI technology affordable and accessible worldwide.

Read More at the original source →