2 Sources
[1]
AI that mimics human problem solving is a big advance, but comes with new risks and problems
by Pasquale Minervini, Edoardo Ponti, Nikolay Malkin , The Conversation OpenAI recently unveiled its latest artificial intelligence (AI) models, o1-preview and o1-mini (also referred to as "Strawberry"), claiming a significant leap in the reasoning capabilities of large language models (the
[2]
AI that mimics human problem solving is a big advance - but comes with new risks and problems
The University of Edinburgh provides funding as a member of The Conversation UK. OpenAI recently unveiled its latest artificial intelligence (AI) models, o1-preview and o1-mini (also referred to as "Strawberry"), claiming a significant leap in the reasoning capabilities of large language models
Share
Copy Link
OpenAI's latest AI models, including "Strawberry," showcase advanced reasoning capabilities but also spark debates about novelty, efficacy, and potential risks in AI development.

OpenAI has recently introduced its latest artificial intelligence models, o1-preview and o1-mini, collectively known as "Strawberry." These models represent a significant advancement in the reasoning capabilities of large language models, the technology underpinning systems like ChatGPT
1
2
.The cornerstone of Strawberry's capabilities is its proficiency in "chain-of-thought reasoning." This approach mirrors human problem-solving methods by breaking down complex tasks into simpler, manageable sub-tasks. It's akin to a person using a notepad to jot down intermediate steps while tackling a problem
1
2
.While chain-of-thought reasoning in AI is not entirely new, its implementation in models not specifically trained for this purpose was first observed in 2022. Research groups, including those led by Jason Wei from Google Research and Takeshi Kojima from the University of Tokyo and Google, pioneered this discovery
1
2
.Earlier contributions to this field include:
Strawberry's potential lies in scaling up these concepts to new heights
1
2
.The exact method employed by OpenAI for Strawberry remains undisclosed. However, experts speculate that it utilizes a procedure known as "self-verification." This process enhances the AI system's ability to perform chain-of-thought reasoning, drawing inspiration from human cognitive processes of reflection and scenario planning
1
2
.Strawberry, like most recent AI systems based on large language models, undergoes a two-stage development:
While Strawberry's self-verification approach is thought to be less data-intensive, there are indications that some o1 models were trained on extensive expert-annotated examples of chain-of-thought reasoning. This raises questions about the balance between self-improvement and expert-guided training in developing its capabilities
1
2
.Related Stories
Despite its advancements, Strawberry still faces limitations:
These factors contribute to potential risks of misinformation and flawed reasoning
1
2
.OpenAI's recent performance evaluation report on o1 models has uncovered some risks:
These findings underscore the need for robust safeguards and ethical considerations in AI development
1
2
.As AI continues to evolve, the balance between technological advancement and responsible implementation remains a critical challenge for researchers, developers, and policymakers alike.
Summarized by
Navi
[1]
1
Technology

2
Technology

3
Policy and Regulation
