This is an archive article published on April 28, 2025
Premium

AI basics | What it means for AI models to ‘reason’, with OpenAI’s ‘smartest’ new o3 and o4-mini models launched

In training the models, OpenAI used “reinforcement learning”. How does this method help AI models make better decisions when replying to user questions?

Through “reasoning”, the AI models consider different approaches and solutions to a prompt, while recognising patterns to arrive at the answer.Through “reasoning”, the AI models consider different approaches and solutions to a prompt, while recognising patterns to arrive at the answer. (Via Freepik)
Written by: Vidhatri Rao
5 min readNew DelhiApr 28, 2025 04:03 PM IST First published on: Apr 28, 2025 at 02:07 PM IST

On April 16, OpenAI released two new Artificial Intelligence (AI) reasoning models named OpenAI o3 and o4-mini, which the company said were the latest “in a series of models trained to think for longer before responding”. The company called them the “smartest models” it has released, “representing a step change in ChatGPT’s capabilities for everyone from curious users to advanced researchers”.

In training both models, the company said it used “reinforcement learning”, a technique previously used by other AI companies, including the Chinese startup DeepSeek. OpenAI has also claimed that, compared to the earlier iterations, its new models should “feel more natural and conversational, especially as they reference memory and past conversations to make responses more personalized and relevant”.

Latest Comment
Post Comment
Read Comments