OpenAI launches new series of AI models with 'reasoning' abilities

Published 09/12/2024, 01:20 PM
Updated 09/13/2024, 01:01 AM
© Reuters. FILE PHOTO: OpenAI logo is seen in this illustration taken May 20, 2024. REUTERS/Dado Ruvic/Illustration/File Photo

By Katie Paul and Anna Tong

(Reuters) -Microsoft-backed OpenAI said on Thursday it was launching its "Strawberry" series of AI models designed to spend more time processing answers to queries in order to solve hard problems.

The models, first reported by Reuters, are capable of reasoning through complex tasks and can solve more challenging problems than previous models in science, coding and math, the AI firm said in a blog post.

OpenAI used the code name Strawberry to refer to the project internally, while it dubbed the models announced on Thursday o1 and o1-mini. The o1 will be available in ChatGPT and its API starting Thursday, the company said.

Noam Brown, a researcher at OpenAI focused on improving reasoning in the company's models, confirmed in a post on social media platform X that the models were the same as the Strawberry project.

"I'm excited to share with you all the fruit of our effort at OpenAI to create AI models capable of truly general reasoning," Brown wrote.

In its blog post, OpenAI said the o1 model scored 83% on the qualifying exam for the International Mathematics Olympiad, compared with 13% for its previous model, GPT-4o.

The model also improved performance on competitive programming questions and exceeded human PhD-level accuracy on a benchmark of science problems, the company said.

Brown said the models were able to accomplish the scores by incorporating a technique known as "chain-of-thought" reasoning, which involves breaking down complex problems into smaller logical steps.

Researchers have noted that AI model performance on complex problems tends to improve when the approach has been used as a prompting technique. OpenAI has now automated this capability so the models can break down problems on their own, without user prompting.

© Reuters. FILE PHOTO: OpenAI logo is seen in this illustration taken May 20, 2024. REUTERS/Dado Ruvic/Illustration/File Photo

"We trained these models to spend more time thinking through problems before they respond, much like a person would. Through training, they learn to refine their thinking process, try different strategies, and recognize their mistakes," OpenAI said.

Reuters was the first to report OpenAI's work on the reasoning project, then called Q*, in November 2023. It reported in July that the project had come to be known as Strawberry.

Latest comments

Risk Disclosure: Trading in financial instruments and/or cryptocurrencies involves high risks including the risk of losing some, or all, of your investment amount, and may not be suitable for all investors. Prices of cryptocurrencies are extremely volatile and may be affected by external factors such as financial, regulatory or political events. Trading on margin increases the financial risks.
Before deciding to trade in financial instrument or cryptocurrencies you should be fully informed of the risks and costs associated with trading the financial markets, carefully consider your investment objectives, level of experience, and risk appetite, and seek professional advice where needed.
Fusion Media would like to remind you that the data contained in this website is not necessarily real-time nor accurate. The data and prices on the website are not necessarily provided by any market or exchange, but may be provided by market makers, and so prices may not be accurate and may differ from the actual price at any given market, meaning prices are indicative and not appropriate for trading purposes. Fusion Media and any provider of the data contained in this website will not accept liability for any loss or damage as a result of your trading, or your reliance on the information contained within this website.
It is prohibited to use, store, reproduce, display, modify, transmit or distribute the data contained in this website without the explicit prior written permission of Fusion Media and/or the data provider. All intellectual property rights are reserved by the providers and/or the exchange providing the data contained in this website.
Fusion Media may be compensated by the advertisers that appear on the website, based on your interaction with the advertisements or advertisers.
© 2007-2025 - Fusion Media Limited. All Rights Reserved.