GPT-4o: What the latest ChatGPT update can do and when you can get it

By Jon Martindale Updated May 14, 2024

OpenAI developer using GPT-4o. — OpenAI

GPT-4o is the latest and greatest large language model (LLM) AI released by OpenAI, and it brings with it heaps of new features for free and paid users alike. It’s a multimodal AI and enhances ChatGPT with faster responses, greater comprehension, and a number of new abilities that will continue to roll out in the weeks to come.

Contents

Availability and price
It’s way faster
Advanced voice support
Improved comprehension
Native macOS desktop app
It’s not all quite ready, yet

With increasing competition from Meta’s Llama 3 and Google’s Gemini, OpenAI’s latest release is looking to stay ahead of the game. Here’s why it’s so exciting.

Availability and price

If you’ve been using the free version of ChatGPT for a while and jealously eyed the features that ChatGPT Plus users have been enjoying, there’s great news! You too can now play around with image detection, file uploads, find custom GPTs in the GPT Store, utilize Memory to retain your conversation as you chat so that you don’t need to repeat yourself, and analyze data and perform complicated calculations.

That’s all alongside the higher intelligence of the standard GPT-4 model, which GPT-4o is an equivalent of, even if it was trained from the ground up as a multimodal AI. The reason this is possible is because GPT-4o is computationally far cheaper to run, meaning it requires fewer tokens, which makes it more viable for a wider user base to enjoy it.

However, free users will have a limited number of messages they can send to GPT-4o per day. When that threshold is reached, you’ll be bumped over to the GPT-3.5 model.

It’s way faster

OpenAI's Mira Murati introduces GPT-4o. — OpenAI

GPT-4 was distinct from GPT-3.5 in a number of ways, and speed was one of them. GPT-4 was just way, way slower, even with its advances in recent months and the introduction of GPT-4 Turbo. However, GPT-4o is almost instantaneous. That makes its text responses far swifter and more actionable, with voice conversations occurring in closer to real- time.

While response speed feels like more of a nice-to-have feature than a game-changing one, the fact that you can get responses in near real time makes GPT-4o a much more viable tool for tasks like translation and conversational help.

Advanced voice support

Although upon its initial debut, GPT-4o is only able to work with text and images, it’s been built from the ground up to utilize voice commands and to be able to interact with users using audio. That means that where GPT-4 could take a voice, convert it into text, respond to that, and then convert its text response to a voice output, GPT-4o can hear a voice, and respond in kind. With its improved speed, it can respond far more conversationally, and can understand unique aspects of voice like tone, pace, mood, and more.

GPT-4o can laugh, be sarcastic, catch itself when making a mistake, and adjust midstream, and you can interrupt it conversationally without that derailing its response. It can also understand different languages and translate on the fly, making it usable as a real-time translation tool. It can sing — or even duet with itself.

Two GPT-4os interacting and singing

This could be used for interview prep, singing coaching, running role-playing NPCs, telling dramatic bedtime stories with different voices and characters, creating voiced dialogue for a game project, telling jokes (and laughing in response to yours), and so much more.

Improved comprehension

GPT-4o understands you much better than its predecessors did, especially if you speak to it. It can read tone and intention far better, and if you want it to be relaxed and friendly, it’ll joke with you in an attempt to keep the conversation light.

When it’s analyzing code or text, it’ll take your intentions into consideration far more, making it better at giving you the response you want and requiring less-specific prompting. It’s better at reading video and images, making it capable of understanding the world around it.

Live demo of GPT-4o vision capabilities

In several demos, OpenAI showed users filming the room they’re in, with GPT-4o models then describing it. In one video, the AI even described the room space to another version of itself, which then had its own responses based on that description.

Native macOS desktop app

The ChatGPT desktop app open in a window next to some code. — OpenAI

Native AI in Windows is still restricted to the very limited Copilot (for now), but macOS users will soon be able to make full use of ChatGPT and its new GPT-4o model right from the desktop. With a new native desktop app, ChatGPT will be more readily available — and with a new user interface to boot — making it easier to use than ever before.

The app will be available for most ChatGPT Plus users in the coming days, and will be rolled out to free users in the coming weeks. A Windows version is promised for later this year.

It’s not all quite ready, yet

At the time of writing, the only aspects of GPT-4o that are available to the public are the text and image modes. There’s no advanced voice support, no real-time video comprehension, and the macOS desktop app won’t be available to everyone for a few more days at least.

But it is all coming. These changes and other exciting upgrades for ChatGPT are just around the corner.

Topics

Evergreen writer

Jon Martindale is a freelance evergreen writer and occasional section coordinator, covering how to guides, best-of lists, and…

Computing

Your ChatGPT conversation history is now searchable

ChatGPT chat search

OpenAI debuted a new way to more efficiently manage your growing ChatGPT chat history on Tuesday: a search function for the web app. With it, you'll be able to quickly surface previous references and chats to cite within your current ChatGPT conversation.

"We’re starting to roll out the ability to search through your chat history on ChatGPT web," the company announced via a post on X (formerly Twitter). "Now you can quickly & easily bring up a chat to reference, or pick up a chat where you left off."

Computing

GPT-5: everything we know so far about OpenAI’s next frontier model

A MacBook Pro on a desk with ChatGPT's website showing on its display.

There's perhaps no product more hotly anticipated in tech right now than GPT-5. Rumors about it have been circulating ever since the release of GPT-4, OpenAI's groundbreaking foundational model that's been the basis of everything the company has launched over the past year, such as GPT-4o, Advanced Voice Mode, and the OpenAI o1-preview.

Those are all interesting in their own right, but a true successor to GPT-4 is still yet to come. Now that it's been over a year a half since GPT-4's release, buzz around a next-gen model has never been stronger.
When will GPT-5 be released?
OpenAI has continued a rapid rate of progress on its LLMs. GPT-4 debuted on March 14, 2023, which came just four months after GPT-3.5 launched alongside ChatGPT. OpenAI has yet to set a specific release date for GPT-5, though rumors have circulated online that the new model could arrive as soon as late 2024.

Computing

The best AI chatbots to try: ChatGPT, Gemini, and more

Bing Chat shown on a laptop.

The idea of chatbots has been around since the early days of the internet. But even compared to popular voice assistants like Siri, the generated chatbots of the modern era are far more powerful.

Yes, you can converse with them in natural language. But these AI chatbots can generate text of all kinds, from poetry to code, and the results really are exciting. ChatGPT remains in the spotlight, but as interest continues to grow, more rivals are popping up to challenge it.
OpenAI ChatGPT and ChatGPT Plus