2 Sources
[1]
Why GPT can't think like us
Artificial Intelligence (AI), particularly large language models like GPT-4, has shown impressive performance on reasoning tasks. But does AI truly understand abstract concepts, or is it just mimicking patterns? A new study from the University of Amsterdam and the Santa Fe Institute reveals that
[2]
Why GPT can't think like us
Artificial Intelligence (AI), particularly large language models like GPT-4, has shown impressive performance on reasoning tasks. But does AI truly understand abstract concepts, or is it just mimicking patterns? A new study from the University of Amsterdam and the Santa Fe Institute reveals that
Share
Copy Link
A new study from the University of Amsterdam and Santa Fe Institute shows that while GPT models perform well on standard analogy tasks, they struggle with variations, indicating limitations in AI's reasoning capabilities compared to humans.

A groundbreaking study conducted by researchers from the University of Amsterdam and the Santa Fe Institute has shed light on the limitations of artificial intelligence (AI) in replicating human-like reasoning. The research, published in Transactions on Machine Learning Research, focused on comparing the performance of GPT models with human cognition in analogical reasoning tasks
1
2
.Analogical reasoning, a fundamental aspect of human cognition, involves drawing comparisons between different concepts based on shared similarities. This ability is crucial for understanding the world and making decisions. For instance, recognizing that "cup is to coffee as soup is to bowl" demonstrates this type of reasoning
1
2
.The study, led by Martha Lewis from the Institute for Logic, Language and Computation at the University of Amsterdam and Melanie Mitchell from the Santa Fe Institute, examined the performance of GPT models and humans on three types of analogy problems. Importantly, the researchers also tested how well both groups handled subtle modifications to these problems
1
2
.While GPT models showed impressive capabilities in solving standard analogy problems, they struggled significantly when faced with variations of these tasks. This contrast was particularly evident in several areas:
Digit Matrices: GPT models' performance dropped noticeably when the position of the missing number was altered, whereas humans had no such difficulty
1
2
.Story Analogies: GPT-4 showed a bias towards selecting the first given answer as correct, a tendency not observed in human participants. The AI also had more trouble than humans when key story elements were reworded
1
2
.Simple Analogy Tasks: On simpler tasks, GPT models' performance declined with modifications, while humans maintained consistent results
1
2
.The research challenges the assumption that AI models like GPT-4 can reason in ways comparable to human cognition. Lewis explains, "This suggests that AI models often reason less flexibly than humans and their reasoning is less about true abstract understanding and more about pattern matching"
1
2
.Related Stories
These findings raise important considerations for the deployment of AI in critical decision-making domains such as education, law, and healthcare. While AI remains a powerful tool, the study emphasizes that it is not yet a suitable replacement for human reasoning and thinking
1
2
.The research underscores the need for continued development in AI to achieve more robust and flexible reasoning capabilities. As AI increasingly integrates into various aspects of society, understanding its limitations and strengths becomes crucial for responsible implementation and development
1
2
.Summarized by
Navi
[1]
[2]
11 Oct 2024•Science and Research

13 May 2025•Science and Research

13 Oct 2024•Science and Research

1
Technology

2
Technology

3
Technology
