Saturday, September 5, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Technology and Engineering

AI makes human-like reasoning mistakes

July 16, 2024
in Technology and Engineering
Denise Maddox
By Denise Maddox Scienmag Editorial Profile - Mechanical Engineering
Reading Time: 3 mins read
0
AI makes human-like reasoning mistakes
67
SHARES
608
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

Large language models (LMs) can complete abstract reasoning tasks, but they are susceptible to many of the same types of mistakes made by humans. Andrew Lampinen, Ishita Dasgupta, and colleagues tested state-of-the-art LMs and humans on three kinds of reasoning tasks: natural language inference, judging the logical validity of syllogisms, and the Wason selection task. The authors found the LMs to be prone to similar content effects as humans. Both humans and LMs are more likely to mistakenly label an invalid argument as valid when the semantic content is sensical and believable. LMs are also just as bad as humans at the Wason selection task, in which the participant is presented with four cards with letters or numbers written on them (e.g., ‘D’, ‘F’, ‘3’, and ‘7’) and asked which cards they would need to flip over to verify the accuracy of a rule such as “if a card has a ‘D’ on one side, then it has a ‘3’ on the other side.” Humans often opt to flip over cards that do not offer any information about the validity of the rule but that test the contrapositive rule. In this example, humans would tend to choose the card labeled ‘3,’ even though the rule does not imply that a card with ‘3’ would have ‘D’ on the reverse. LMs make this and other errors but show a similar overall error rate to humans. Human and LM performance on the Wason selection task improves if the rules about arbitrary letters and numbers are replaced with socially relevant relationships, such as people’s ages and whether a person is drinking alcohol or soda. According to the authors, LMs trained on human data seem to exhibit some human foibles in terms of reasoning—and, like humans, may require formal training to improve their logical reasoning performance.

Reasoning test expamples

Credit: Lampinen et al

Large language models (LMs) can complete abstract reasoning tasks, but they are susceptible to many of the same types of mistakes made by humans. Andrew Lampinen, Ishita Dasgupta, and colleagues tested state-of-the-art LMs and humans on three kinds of reasoning tasks: natural language inference, judging the logical validity of syllogisms, and the Wason selection task. The authors found the LMs to be prone to similar content effects as humans. Both humans and LMs are more likely to mistakenly label an invalid argument as valid when the semantic content is sensical and believable. LMs are also just as bad as humans at the Wason selection task, in which the participant is presented with four cards with letters or numbers written on them (e.g., ‘D’, ‘F’, ‘3’, and ‘7’) and asked which cards they would need to flip over to verify the accuracy of a rule such as “if a card has a ‘D’ on one side, then it has a ‘3’ on the other side.” Humans often opt to flip over cards that do not offer any information about the validity of the rule but that test the contrapositive rule. In this example, humans would tend to choose the card labeled ‘3,’ even though the rule does not imply that a card with ‘3’ would have ‘D’ on the reverse. LMs make this and other errors but show a similar overall error rate to humans. Human and LM performance on the Wason selection task improves if the rules about arbitrary letters and numbers are replaced with socially relevant relationships, such as people’s ages and whether a person is drinking alcohol or soda. According to the authors, LMs trained on human data seem to exhibit some human foibles in terms of reasoning—and, like humans, may require formal training to improve their logical reasoning performance.



Journal

PNAS Nexus

Article Title

Language models, like humans, show content effects on reasoning tasks

Article Publication Date

16-Jul-2024

COI Statement

All authors are employed by Google DeepMind; J.L.M. is affiliated part-time.

Subject of Research: Technology and Engineering

Article Title: AI makes human-like reasoning mistakes

Article References: Original research article

Image Credits: AI Generated

DOI: Not provided

Keywords: Not provided

Cite Scienmag News

Denise Maddox. (July 16, 2024). AI makes human-like reasoning mistakes. Scienmag. https://scienmag.com/ai-makes-human-like-reasoning-mistakes/

Denise Maddox. "AI makes human-like reasoning mistakes." Scienmag, 16 July 2024, https://scienmag.com/ai-makes-human-like-reasoning-mistakes/. Accessed 5 September 2026.

Denise Maddox. "AI makes human-like reasoning mistakes." Scienmag. July 16, 2024. https://scienmag.com/ai-makes-human-like-reasoning-mistakes/

Share27Tweet17
Previous Post

Ultrasonography of hepatocellular carcinoma: From diagnosis to prognosis

Next Post

Should AI be used in psychological research?

Related Posts

Kagome metals enable goniopolar transverse thermoelectric effects via Fermiology
Technology and Engineering

Kagome metals enable goniopolar transverse thermoelectric effects via Fermiology

September 5, 2026
Exciton condensate defect-bound states mirror Yu-Shiba-Rusinov physics
Technology and Engineering

Exciton condensate defect-bound states mirror Yu-Shiba-Rusinov physics

September 5, 2026
Measuring elastic barriers that block molecular glass rearrangements
Technology and Engineering

Measuring elastic barriers that block molecular glass rearrangements

September 5, 2026
Retinomorphic sensor adapts optoelectronic computing through multiple stimulus responses
Technology and Engineering

Retinomorphic sensor adapts optoelectronic computing through multiple stimulus responses

September 5, 2026
CNN architectures advance crop disease detection and severity quantification
Technology and Engineering

CNN architectures advance crop disease detection and severity quantification

September 5, 2026
Fast Semi-Supervised Node Embeddings Using Structural and Label Optimization
Technology and Engineering

Fast Semi-Supervised Node Embeddings Using Structural and Label Optimization

September 5, 2026
Next Post
Should AI be used in psychological research?

Should AI be used in psychological research?

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • Organoid co-cultures expose epithelial and fibroblast diversity in pancreatic cancer and pancreatitis
  • Road expansion has quadrupled fragmentation of Siberian intact forests since 2000
  • Genome-wide study across ancestries reveals genetic roots of Hashimoto’s thyroiditis
  • Pregnancy haemoglobin levels linked to maternal and newborn health outcomes

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,151 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading