Monday, October 5, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Psychology & Psychiatry

AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function

October 5, 2026
in Psychology & Psychiatry
Glenn Wilkins
By Glenn Wilkins Scienmag Editorial Profile - Clinical Psychology
Reading Time: 5 mins read
0
AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function

AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

Executive functioning — the umbrella term for the mental machinery behind working memory, cognitive flexibility, and inhibitory control — shapes nearly everything we do, from solving a math problem to resisting the pull of a smartphone notification. Yet measuring it has always been slow, noisy, and inefficient. A new study published in Behavior Research Methods by Robert Kasumba, Dennis L. Barbour, and colleagues at Washington University in St. Louis and their collaborators now shows, through rigorously controlled simulations, that a pair of machine learning techniques can estimate a person’s executive function profile with a fraction of the data that conventional testing demands. The findings could reshape how psychologists, clinicians, and educators assess cognition, potentially turning hour-long test batteries into brief, adaptive sessions tailored to each individual brain.

The core problem with traditional cognitive assessment is statistical. Standard batteries such as the Corsi Block-Tapping task for working memory or the Stroop test for inhibitory control treat each task as an isolated measurement. Researchers typically average across repeated trials — mean response times on the Stroop, for example — and treat trial-to-trial variability as mere noise. Structural equation models can link tasks together, but only under assumptions of linearity that the authors argue oversimplify the true architecture of cognition. The result is a testing paradigm that requires many observations per task before estimates stabilize, and one that fails entirely when a task is skipped or data are missing.

The Washington University team had previously introduced an alternative: the distributional latent variable model, or DLVM. Instead of summarizing performance with a single average, DLVM models the full distribution of an individual’s behavior on each task — capturing both central tendency and variability — and compresses that information into a low-dimensional latent space learned by a neural network. Crucially, every single observation contributes to the estimation of multiple constructs at once, because the model exploits the nonlinear dependencies that link performance across tasks. A person’s position in this learned latent space constitutes their cognitive profile, and the trained model can even generate plausible behavioral data for hypothetical profiles, functioning as a generative oracle for simulation.

Paired with DLVM is a second innovation: distributional active learning, or DALE, a Bayesian algorithm that decides which test item to administer next. DALE frames cognitive testing as sequential Bayesian inference. After each response, it updates a posterior distribution over the individual’s latent position and then selects the next trial by maximizing expected mutual information — in plain terms, it always asks the question that will do the most to shrink its own uncertainty. The algorithm can be primed with a small batch of observations spanning all tasks, in this case two samples per task, before active selection kicks in. This lineage descends from Bayesian active learning methods that have already transformed audiology and vision testing, where adaptive stimulus selection dramatically reduced the trials needed to map perceptual thresholds.

What has been missing until now is ground truth. In the team’s earlier human study, real participants produced real data, but the true underlying cognitive parameters were unknown, making it impossible to measure estimation accuracy precisely. The new paper solves this with an elegant simulation strategy. The researchers trained DLVM models on a retrospective dataset — 88 valid testing sessions in which 18 participants completed up to ten sessions of an eight-task battery over ten days via a mobile app, covering tasks such as the Paced Auditory Serial Addition Test, Countermanding, Running Span, Numerical Stroop, Magnitude Comparison, Corsi Simple and Complex Span, and Cancellation. They then sampled 88 points systematically across the learned latent space and used the model to generate the corresponding ground-truth distributional parameters, from which 240 trial-level observations per task were simulated. Every estimate could now be checked against a known answer.

The first set of analyses pitted DLVM against independent maximum likelihood estimation, or IMLE, the optimal approach if tasks truly were independent. Both models received identical data under equal allocation. The verdict was striking. With only two observations per task, DLVM with two latent dimensions achieved Kullback–Leibler divergence values below 0.200 across all tasks, consistently beating IMLE, with the biggest gains on the sigmoid-shaped span tasks that are notoriously data-hungry. DLVM held its advantage until roughly seven observations per task. More dramatically, in validation analyses DLVM needed only about 20 observations per task — 160 total — to reach near-perfect accuracy in recovering marginal distributions, whereas IMLE required about 100 per task, or 800 total. DLVM could even estimate parameters for tasks that were never administered, something IMLE fundamentally cannot do.

The second stage asked how the way data are collected changes the picture. The team compared six configurations: DLVM or IMLE, each fed by DALE’s adaptive sampling, uniform random sampling, or a traditional fixed test battery delivered in sequential blocks. DALE combined with DLVM was the clear winner in the sparse-data regime, driving KLD below 0.05 by roughly 80 observations. The adaptive algorithm concentrated trials on the complex distributional tasks that carried the most information while allocating fewer trials to simpler accuracy-based tasks, and each simulated session received its own unique, personalized battery. The contrast with conventional practice was stark: at 80 total observations, when a fixed battery had covered only three tasks, DLVM with the battery achieved a KLD of 0.148 while IMLE with the same battery sat at a catastrophic 39.8, simply because it could not infer anything about tasks it had not yet reached.

The authors were careful to probe their own assumptions. Because the simulated data were generated from a DLVM-learned latent space, DLVM might enjoy a structural home-field advantage. So they repeated the analysis using an IMLE-based generative process instead. The qualitative pattern held: IMLE performed best when recovering parameters from data generated under its own specification with abundant data, but DLVM and DALE retained their advantage whenever data were sparse. The team also examined DALE’s trajectories through latent space, finding that the algorithm made large corrections in the first trials, converged to a localized region after about 30 observations, and reliably landed in regions of high probability — even when those regions did not coincide exactly with the true latent position, a consequence of the nonlinear latent space admitting multiple equally plausible solutions. Mean root mean squared error across all 88 sessions was 1.02, and only seven sessions converged to positions with normalized negative log probability above 0.05.

Perhaps the most counterintuitive insight concerns the trade-off between model flexibility and data hunger. In most of machine learning, highly flexible models like deep neural networks need enormous datasets to converge. Here the logic inverts: IMLE, the more flexible estimator, only overtakes DLVM once more than 800 observations are available under these testing conditions, while DLVM’s deliberately constrained low-dimensional embedding extracts meaningful inference from a handful of trials. The authors even observed that random sampling eventually surpassed active learning at very large sample counts, suggesting that switching to random sampling once DALE plateaus — or refining its acquisition function — could reveal additional structure in the data. Both of the study’s pre-registered hypotheses were supported by the results.

The practical implications are considerable. DALE could support adaptive assessment in clinical screening, longitudinal monitoring of cognitive change, and large-scale educational studies where testing time is limited and participant burden matters. One examinee might receive extra Stroop trials while another gets more span or countermanding items, depending on where their performance remains most uncertain, and the output is an uncertainty-aware summary of observable performance rather than a brittle point estimate. The authors argue that future behavioral tasks should be designed to be multidimensional and fully featured so that adaptive algorithms never need to repeat an item, since unsampled regions of a feature space are almost always more informative than sampled ones. For a field long anchored to rigid, hour-long batteries, the message is clear: cognition can be measured faster, smarter, and more personally than ever before.

Subject of Research: Bayesian distributional latent variable modeling and adaptive active learning for efficient assessment of executive functioning

Article Title: Bayesian distributional models of executive functioning

Article References: Bayesian distributional models of executive functioning. (n.d.). https://doi.org/10.3758/s13428-026-03191-x

Image Credits: AI Generated

DOI: 10.3758/s13428-026-03191-x

Keywords: executive function, Bayesian inference, active learning, latent variable models, machine learning, cognitive assessment, working memory, inhibitory control, cognitive flexibility, psychometrics, neural networks, adaptive testing

Cite Scienmag News

Glenn Wilkins. (October 5, 2026). AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function. Scienmag. https://scienmag.com/ai-powered-bayesian-models-slash-the-time-needed-to-measure-executive-function/

Glenn Wilkins. "AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function." Scienmag, 5 October 2026, https://scienmag.com/ai-powered-bayesian-models-slash-the-time-needed-to-measure-executive-function/. Accessed 5 October 2026.

Glenn Wilkins. "AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function." Scienmag. October 5, 2026. https://scienmag.com/ai-powered-bayesian-models-slash-the-time-needed-to-measure-executive-function/

Tags: active learningadaptive cognitive testing methodsadaptive testingAI-driven cognitive measurementBayesian inferenceBayesian models for cognitive testingcognitive assessmentcognitive flexibilityestimating executive function with minimal dataExecutive functionexecutive function assessmentimproving efficiency in mental ability assessmentsinhibitory controlinnovative approaches to working memory and inhibitory controllatent variable modelsMachine learningmachine learning in psychologyneural networkspersonalized neuropsychological testingpsychometricsrapid assessment of executive functioningsimulation-based validation of cognitive modelsstatistical methods in cognitive neuroscienceworking memory
Share26Tweet16
Previous Post

Malawi’s transparency law looks strong on paper but hides a design flaw

Next Post

Soil pH Emerges as Master Switch Governing Legume Farming Performance Across China

Related Posts

Feeling Good Pays Off: How Positive Emotions Quietly Drive University Success
Psychology & Psychiatry

Feeling Good Pays Off: How Positive Emotions Quietly Drive University Success

October 5, 2026
How US Perinatal Psychiatry Access Programs Are Training Frontline Clinicians to Tackle Maternal Mental Health
Psychology & Psychiatry

How US Perinatal Psychiatry Access Programs Are Training Frontline Clinicians to Tackle Maternal Mental Health

October 5, 2026
Women With High Blood Pressure Face Twice the Odds of Suicidal Thoughts, Three-Country Study Finds
Psychology & Psychiatry

Women With High Blood Pressure Face Twice the Odds of Suicidal Thoughts, Three-Country Study Finds

October 5, 2026
Unpredictable Childhoods May Fuel Problematic Pornography Use in University Students, Study Finds
Psychology & Psychiatry

Unpredictable Childhoods May Fuel Problematic Pornography Use in University Students, Study Finds

October 5, 2026
Psychological Distress, Not Personality, Drives the Body’s Complaints of Somatization
Psychology & Psychiatry

Psychological Distress, Not Personality, Drives the Body’s Complaints of Somatization

October 5, 2026
New Bedside Cognitive Test Tracks Brain Fog After Electroconvulsive Therapy
Psychology & Psychiatry

New Bedside Cognitive Test Tracks Brain Fog After Electroconvulsive Therapy

October 5, 2026
Next Post
Soil pH Emerges as Master Switch Governing Legume Farming Performance Across China

Soil pH Emerges as Master Switch Governing Legume Farming Performance Across China

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • How Sugar Rules the Tea Plant: New Review Reveals Carbon Secrets of the World’s Favorite Drink
  • Soil pH Emerges as Master Switch Governing Legume Farming Performance Across China
  • AI-Powered Bayesian Models Slash the Time Needed to Measure Executive Function
  • Malawi’s transparency law looks strong on paper but hides a design flaw

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,150 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading