Spotify hit band The Velvet Sundown comes clean on AI
The Velvet Sundown burst onto the music scene in early June and in the space of just a few...
Although the entire AI boom was triggered by just one ChatGPT model, a lot has changed since 2022. New models have been released, old models have been replaced, updates roll out and roll back again when they go wrong â the world of LLMs is pretty busy. At the moment, we have six OpenAI LLMs to choose from and, as both users and Sam Altman are aware, their names are completely useless.
Most people have probably just been using the newest model they can get their hands on, but it turns out that each of the six current models is good at different things â and OpenAI has finally decided to tell us which model to use for which tasks.
LLMs are unpredictable â users never know what kind of responses they will get, and the developers donât really know either. Sure, it might be more convenient if we had all of the capabilities available rolled up into one model, but that isnât as easy as it sounds.
As OpenAI tweaks its models, some things get better and other things get worse â and sometimes unexpected side effects occur. Thereâs no telling how long it would take to balance things out perfectly, so it makes more sense to just release new versions even when improvements are only focused on a few areas.
The results of this approach are the six main models we have right now: GPT-4o, GPT-4.5, OpenAI o4-mini, OpenAI o4-mini-high, OpenAI o3, and OpenAI o1 pro mode. And Iâm just going to say it again â these names really are useless. OpenAI may have given us a document explaining what each one does now, but that doesnât mean youâll be able to remember which name matches which capabilities â so consider saving this little cheat sheet from the document if you need to remember.
Part of the latest 4o family of models, GPT-4o âexcels at everyday tasks.â This includes:
You can search the web with it, generate images, use advanced voice features, analyze data, and create custom GPTs. You can also upload various file types to aid your prompts.
According to OpenAIâs own research, however, 4o does have a bit of a hallucination problem. Itâs not the worst of the bunch, but it did hallucinate around twice as much as o1 during testing.
This can be problematic if youâre using it to search the web or learn new things â the trickiest aspect of hallucinations is that they often sound entirely plausible, making it harder to just âcheck when something sounds off.â Instead, the only way to be sure is to check just about everything that you donât already know to be true.
According to OpenAI, GPT-4.5âs strong suit is emotional intelligence. This means it should be good at helping you communicate with other people, with official recommendations including:
With other strengths such as clear communication and creativity, GPT-4.5 is better equipped to help you find the perfect tone or phrasing for specific situations â and make sure everything still sounds human.
One of the more terribly named models, o4-mini drops the âGPTâ element of the naming scheme and awkwardly swaps the 4o around to o4. Itâs a smaller model, which means itâs not stuffed to the brim with as much random internet information as a full-sized model.
The upside of this is that itâs quick and less expensive to run, and the downside is that the model has less âworld knowledgeâ and is prone to hallucinating to make up for that.
Instead of asking it questions about the world, OpenAI recommends using o4-mini for fast technical tasks. Examples include:
Hereâs another terrible name when viewed in isolation, but fairly easy to understand if you already know what OpenAI o4-mini is. Itâs still a small model, but itâs a step up from the normal o4-mini because it âthinks longer for higher accuracy.â
This makes it better at more detailed coding tasks, math, and scientific explanations. Here are OpenAIâs examples:
This is technically an older model (because it doesnât have a â4â), but because the o4/4o family didnât make improvements in every area, itâs still very relevant. o3 is particularly good at complex, multi-step tasks â the kind of projects that need to be done in multiple stages with multiple prompts.
This includes strategic planning, detailed analyses, extensive coding, advanced math, science, and visual reasoning. If you want to start a task that you know will take a multiple-prompt session to finish, using o3 will help minimize the chances of the model losing track of the context or hallucinating halfway through.
OpenAI suggests use cases like:
OpenAI o1 is now considered a âlegacy model,â though it isnât even a year old yet. The âpro modeâ version is tuned for complex reasoning â which means it takes more time to think, but in return gives better thought-out responses.
o1 also gets the best scores on OpenAIâs PersonQA evaluation, which measures the rate of hallucination. During testing, o1 hallucinates around half as much as o3 and three times less than smaller models like 04-mini. If youâre a big ChatGPT user and your sessions tend to run long, then minimizing the rate of hallucinations could save you a decent chunk of time in the long run.
Here are OpenAIâs examples:
Unfortunately, you can only access GPT-4o and GPT-4o mini on OpenAIâs free tier. If youâre a Plus, Pro, Team, or Enterprise user, you can use the model selector to choose which model you want to use.
ChatGPT is also integrated into various other third-party products, both free and paid, so itâs worth checking which models different products use. For example, my paid search engine, Kagi, gives me access to multiple OpenAI models. There are also lots of other AI aggregate services out there that give you access to multiple models from OpenAI and other companies for a more affordable price than subscribing to each company separately.
While this information about the different models is useful to have, it doesnât affect everyone. If you mostly use ChatGPT to generate images, search the web, and send general queries, then the default GPT-4o is totally fine. Itâs only if youâre into programming, math, science, or particularly large projects that you might want to think about which model is best for the job.
The Velvet Sundown burst onto the music scene in early June and in the space of just a few...
With the launch of GPT-5, ChatGPT received more than just an upgrade â it appears to hav...
After months of teasers, previews, and select rollouts, Microsoftâs Copilot Vision is no...
If you are an Android user wanting to avoid Google Gemini as your default digital assistan...
Windows 11 has support for voice commands like âOpen Edgeâ largely for accessibility p...
OpenAIFollowing its U.S. debut in January, OpenAIâs Operator AI agent will soon be exp...
Comments on "ChatGPT models explained: How to use each, according to OpenAI" :