While GPT-4 performs well in structured reasoning tasks, a new study shows that its ability to adapt to variations is weak—suggesting AI still lacks true abstract understanding and flexibility in decision-making.
Artificial Intelligence (AI), particularly large language models like GPT-4, has shown impressive performance on reasoning tasks. But does AI truly understand abstract concepts, or is it just mimicking patterns? A new study from the University of Amsterdam and the Santa Fe Institute reveals that while GPT models perform well on some analogy tasks, they fall short when the problems are altered, highlighting key weaknesses in AI’s reasoning capabilities.
Analogical reasoning is the ability to draw a comparison between two different things based on their similarities in certain aspects. It is one of the most common methods by which human beings try to understand the world and make decisions. An example of analogical reasoning: cup is to coffee as soup is to ??? (the answer being: bowl)
Large language models like GPT-4 perform well on various tests, including those requiring analogical reasoning. But can AI models truly engage in general, robust reasoning, or do they over-rely on patterns from their training data? This study by language and AI experts Martha Lewis (Institute for Logic, Language and Computation at the University of Amsterdam) and Melanie Mitchell (Santa Fe Institute) examined whether GPT models are as flexible and robust as humans in making analogies. ‘This is crucial, as AI is increasingly used for decision-making and problem-solving in the real world,’ explains Lewis.
Comparing AI models to human performance
Lewis and Mitchell compared the performance of humans and GPT models on three different types of analogy problems:
- Letter sequences – Identify patterns in letter sequences and complete them correctly.
- Digit matrices – Analyzing number patterns and determining the missing numbers.
- Story analogies – Understanding which of two stories best corresponds to a given example story.
A system that truly understands analogies should maintain high performance even on variations
In addition to testing whether GPT models could solve the original problems, the study examined how well they performed when the problems were subtly modified. ‘A system that truly understands analogies should maintain high performance even on these variations’, state the authors in their article.
GPT models struggle with robustness
Humans maintained high performance on most modified versions of the problems, but GPT models, while performing well on standard analogy problems, struggled with variations. ‘This suggests that AI models often reason less flexibly than humans, and their reasoning is less about true abstract understanding and more about pattern matching,’ explains Lewis.
In digit matrices, GPT models showed a significant performance drop when the missing number’s position changed. Humans had no difficulty with this. In story analogies, GPT-4 tended to select the first given answer as correct more often, whereas humans were not influenced by answer order. Additionally, GPT-4 struggled more than humans when key elements of a story were reworded, suggesting a reliance on surface-level similarities rather than deeper causal reasoning.
When tested on modified versions, GPT models showed a decline in performance on simpler analogy tasks, while humans remained consistent. However, both humans and AI struggled with more complex analogical reasoning tasks.
Weaker than human cognition
This research challenges the widespread assumption that AI models like GPT-4 can reason in the same way humans do. ‘While AI models demonstrate impressive capabilities, this does not mean they truly understand what they are doing,’ conclude Lewis and Mitchell. ‘Their ability to generalize across variations is still significantly weaker than human cognition. GPT models often rely on superficial patterns rather than deep comprehension.’
This is a critical warning about using AI in important decision-making areas such as education, law, and healthcare. While AI can be a powerful tool, it is not yet a replacement for human thinking and reasoning.
- Lewis, Martha, and Melanie Mitchell. “Evaluating the Robustness of Analogical Reasoning in Large Language Models.” Transactions on Machine Learning Research, 2025, openreview.net/forum?id=t5cy5v9wp
News
Why More People in Their 30s Are Suddenly Getting Colon Cancer
A major Swiss study found that colorectal cancer is becoming increasingly common in adults under 50, even as rates decline in older age groups. Researchers in Switzerland have identified a concerning trend: while colorectal [...]
Researchers Compare MS Models to Human Tissue in Search for Better Therapies
Researchers identified key differences between two widely used multiple sclerosis models, showing how each can better study myelin damage, immune responses, and repair. The findings may improve efforts to develop treatments that restore lost [...]
Scientists Discover Genetic “Off Switch” That Supercharges CAR T Cells Against Cancer
A new study reveals a possible way to make CAR T-cell therapy more durable and effective by targeting a single gene-regulating protein. CAR T-cell therapy is widely seen as a breakthrough in personalized cancer [...]
New Vitamin B12-Based Therapy Could Change How Brain Cancer Is Treated
Researchers have identified a vitamin B12–based compound that appears capable of crossing the blood–brain barrier and selectively accumulating in glioblastoma tissue. For decades, one of the biggest problems in brain cancer treatment has had [...]
Simple Fiber Supplement Cuts Knee Arthritis Pain in Just 6 Weeks, Study Finds
A daily inulin supplement may help reduce knee osteoarthritis pain while revealing a possible link between gut health, muscle function, and pain sensitivity. For millions of people living with knee osteoarthritis, managing chronic pain [...]
This Common Vitamin May Help Stop Prediabetes From Turning Into Diabetes
Vitamin D may help prevent type 2 diabetes in people with specific genetic variations, offering a possible path toward personalized diabetes prevention. More than 40% of U.S. adults have prediabetes, a condition in which [...]
Ebola, hantavirus: Is the world prepared for the next pandemic?
Funding cuts to health research and a growing antivaccine movement are making it harder than ever to respond to viruses. The World Health Organization (WHO) has declared that an Ebola outbreak in Uganda and [...]
May 2026 Healthcare News and Trends: Market Signals That Matter
Artificial intelligence is dominating headlines, telehealth has settled into a new normal, and digital health continues to promise transformation. However, much of what is being discussed in healthcare today reflects potential rather than reality. [...]
Scientists Rewire Donor Stem Cells To Outsmart Aggressive Blood Cancers
Researchers have tested a gene-edited stem cell transplant designed to shield healthy blood-forming cells from powerful cancer-targeting immunotherapies. For patients with highly aggressive blood cancers, stem cell transplantation can offer a rare chance at [...]
Recent Digital Health Trends, Insights and News – May 2026
Last month marked continued progress as digital health moves into its next phase — from AI expanding into drug discovery and core infrastructure to new federal pathways accelerating device access and home-based care. Together, [...]
Cancer Mystery Solved: Scientists Discover How Melanoma Becomes “Immortal”
Scientists have uncovered a previously overlooked mechanism that may help melanoma cells become effectively “immortal.” Cancer cells face a major problem before they can become deadly: They have to figure out how to stop [...]
How Visual Neurons Organize Thousands of Synaptic Inputs
Summary: A new study uncovered the organizational rules that determine how neurons in the primary visual cortex process information. By imaging both the cell bodies (soma) and the individual synapses (on dendritic spines) of [...]
Scientists Just Found a Surprising Way To Destroy “Forever Chemicals”
Scientists have uncovered a new mechanism that may help break down highly persistent PFAS pollutants. PFAS have earned the nickname “forever chemicals” for a reason. These industrial compounds are so chemically durable that they [...]
Scientists Discover Cheap Material That Kills Deadly Superbugs
A new sulfur-rich antimicrobial polymer shows strong effectiveness against fungal and bacterial pathogens and may offer an affordable solution to antimicrobial resistance. Antimicrobial resistance is creating growing challenges for both healthcare and food production, [...]
What to Know About Cicada, or BA.3.2, the Latest SARS-CoV-2 Variant Under Monitoring
Like periodical cicadas, the insects for which it is nicknamed, SARS-CoV-2 Omicron subvariant BA.3.2 is only just beginning to emerge after lying low for an extended period since it first appeared. Although it was [...]
Scientists Say This Simple Supplement May Actually Reverse Heart Disease
Scientists in Japan say a common supplement may actually help “unclog” certain diseased heart arteries from the inside out. A simple food supplement sold in Japan may have helped reverse a dangerous form of [...]















