Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

When OpenAI introduced GPT-4o in May 2024, it presented the model as offering GPT-4-level intelligence with faster performance and improvements across text, voice, and vision. Its five clearest differences were its multimodal design, more responsive voice interaction, stronger image understanding, language improvements, and launch-era speed and access claims. These are historical launch comparisons—not current ChatGPT options: OpenAI later retired both GPT-4 and GPT-4o from ChatGPT.

What changed between GPT-4o and GPT-4?

GPT-4o (“o” for “omni”) was announced as a model designed to handle text, vision, and audio together, rather than treating voice as a separate, slower chain of processing steps. OpenAI described it as matching GPT-4-level intelligence while operating faster and improving capabilities across text, voice, and vision. That was OpenAI’s launch positioning, not a single independent head-to-head test proving GPT-4o won every comparison.

There is an important timing distinction: GPT-4o’s announced design included more capabilities than were available in ChatGPT on launch day. The initial rollout covered text and images; OpenAI said new audio and video capabilities would arrive later. OpenAI’s May 2024 ChatGPT announcement and GPT-4o overview describe the launch claims and planned capabilities.

1. Multimodal design: text, vision, and audio in one model

OpenAI described GPT-4o as trained end-to-end across text, vision, and audio, with inputs and outputs handled by the same neural network. Its overview says the model can accept combinations of text, audio, images, and video, and generate text, audio, and images. The central upgrade was therefore not just adding a voice feature: it was designing the model to work across modalities as part of one system.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

That description is about the model’s design and announced capabilities, not a promise that every input and output type was already available to every ChatGPT user in May 2024. The launch began with text and image capabilities in ChatGPT; audio and video features were described as forthcoming.

2. Voice interaction: lower reported response latency

OpenAI reported that GPT-4o could respond to audio inputs in as little as 232 milliseconds, with an average of 320 milliseconds. In the same announcement, it said the previous ChatGPT Voice Mode averaged 5.4 seconds of latency with GPT-4. Those numbers describe OpenAI’s reported experience with its prior voice pipeline and the new model—not a universal, controlled comparison of every GPT-4 and GPT-4o interaction.

The practical aim was a more natural conversation: less waiting for a response and better support for speech-to-speech interaction. OpenAI’s GPT-4o overview discusses the latency figures and the model’s audio capabilities. Its GPT-4o system card also covers evaluation and mitigations for text, image, and audio risks; it does not establish that GPT-4o was simply safer overall than GPT-4.

3. Vision: more capable discussion of shared images

OpenAI said GPT-4o was much better at understanding and discussing images shared with it. Its launch example was practical: photograph a menu in another language, ask for a translation, then ask about the food’s history and meaning. This illustrates how image input could support a multi-step conversation, rather than stopping at a basic description.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

This is an improvement OpenAI claimed in its launch announcement, not an independently verified result in a controlled comparison with the original GPT-4. The announcement also should not be read as saying every image-related feature was available to every user at launch.

4. Language performance: quality and speed improvements

OpenAI said GPT-4o improved language capabilities in both quality and speed, and highlighted better non-English text performance compared with GPT-4 Turbo. That specific baseline matters: the announcement does not establish a formal measured language-performance victory over the original GPT-4 across all languages or tasks.

For readers choosing between the names in older comparisons, the safe interpretation is that GPT-4o was pitched as a faster, improved language model, with a particular non-English comparison to GPT-4 Turbo—not that every use of GPT-4o would produce a better answer than every use of GPT-4.

5. Speed and access: distinct launch claims for API and ChatGPT

OpenAI’s May 2024 API announcement said GPT-4o was twice as fast, half the price, and had five times higher rate limits than GPT-4 Turbo. These are API comparisons against GPT-4 Turbo, not figures comparing GPT-4o with the original GPT-4 in ChatGPT. They are historical launch claims and should not be treated as current API pricing, speed, or limits.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

ChatGPT access was described separately. OpenAI said the initial rollout covered Free, Plus, and Team, with Enterprise availability planned. It also said Plus users would receive up to five times the Free plan’s message limits. Those were May 2024 rollout details, not present-day plan limits or availability. See OpenAI’s launch announcement and API overview for the original context.

Independent reader supportYour contribution helps us test, update, and keep practical guides available for everyone.Support on Ko-Fi

Can you still select GPT-4 or GPT-4o in ChatGPT?

No. OpenAI’s release notes say GPT-4 was retired from ChatGPT effective April 30, 2025, and that GPT-4o has since also been retired from ChatGPT. Neither should be described as a currently selectable ChatGPT model. OpenAI’s 2025 GPT-4 retirement notice said GPT-4 would remain available in the API at that time, but that notice does not establish present-day API availability, model identifiers, pricing, or limits.

For current ChatGPT model availability, consult OpenAI’s rolling ChatGPT release notes. The May 2024 comparison remains useful for understanding what GPT-4o introduced, but its launch-era access and performance claims are not a guide to today’s plans or model lineup.

How to read the comparison fairly

  • Separate design from availability: GPT-4o was announced as multimodal, but ChatGPT’s initial rollout did not include every planned audio and video capability.
  • Check the baseline: OpenAI’s language claim highlighted GPT-4 Turbo; its API speed, price, and rate-limit figures also compared GPT-4o with GPT-4 Turbo.
  • Keep figures tied to their setup: the voice-latency numbers refer to OpenAI’s reported audio and prior Voice Mode experience, not every conversation.
  • Treat access and prices as historical: the Free, Plus, Team, Enterprise, and API statements describe May 2024 rollout plans and claims, not current availability.

Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.

Free tools Windows power users keep installed

One-click scans. No signup required.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.