<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>Short Grit Scale &#8211; Science</title>
	<atom:link href="https://scienmag.com/tag/short-grit-scale/feed/" rel="self" type="application/rss+xml" />
	<link>https://scienmag.com</link>
	<description></description>
	<lastBuildDate>Sat, 26 Sep 2026 00:39:29 +0000</lastBuildDate>
	<language>en-US</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	<generator>https://wordpress.org/?v=7.1.2</generator>

<image>
	<url>https://scienmag.com/wp-content/uploads/2024/07/cropped-scienmag_ico-32x32.jpg</url>
	<title>Short Grit Scale &#8211; Science</title>
	<link>https://scienmag.com</link>
	<width>32</width>
	<height>32</height>
</image> 
<site xmlns="com-wordpress:feed-additions:1">73899611</site>	<item>
		<title>Hidden Wording Effects Distort Psychology Scores, But a New Statistical Fix Emerges</title>
		<link>https://scienmag.com/hidden-wording-effects-distort-psychology-scores-but-a-new-statistical-fix-emerges/</link>
		
		<dc:creator><![CDATA[Glenn Wilkins]]></dc:creator>
		<pubDate>Sat, 26 Sep 2026 00:39:29 +0000</pubDate>
				<category><![CDATA[Psychology & Psychiatry]]></category>
		<category><![CDATA[confirmatory factor analysis]]></category>
		<category><![CDATA[dimensionality]]></category>
		<category><![CDATA[enhancing validity of psychological scales]]></category>
		<category><![CDATA[Exploratory Graph Analysis]]></category>
		<category><![CDATA[factor analysis]]></category>
		<category><![CDATA[impact of question phrasing on survey responses]]></category>
		<category><![CDATA[improving psychological assessment accuracy]]></category>
		<category><![CDATA[influence of positive and negative item wording]]></category>
		<category><![CDATA[method variance]]></category>
		<category><![CDATA[parallel analysis]]></category>
		<category><![CDATA[psychological measurement]]></category>
		<category><![CDATA[psychometric method variance]]></category>
		<category><![CDATA[psychometrics]]></category>
		<category><![CDATA[questionnaire bias correction methods]]></category>
		<category><![CDATA[random intercept item factor analysis]]></category>
		<category><![CDATA[randomized intercept item factor analysis]]></category>
		<category><![CDATA[research on survey response distortions]]></category>
		<category><![CDATA[Short Grit Scale]]></category>
		<category><![CDATA[statistical techniques in psychometrics]]></category>
		<category><![CDATA[structural validity]]></category>
		<category><![CDATA[survey design bias]]></category>
		<category><![CDATA[survey methodology]]></category>
		<category><![CDATA[wording effect in psychological questionnaires]]></category>
		<category><![CDATA[wording effects]]></category>
		<guid isPermaLink="false">https://scienmag.com/?p=215687</guid>

					<description><![CDATA[A new study in Behavior Research Methods shows that random intercept item factor analysis can separate genuine trait variance from wording artifacts in mixed-worded psychological scales, using the Short Grit Scale as a test case.]]></description>
										<content:encoded><![CDATA[<p>Every year, millions of people fill out psychological questionnaires designed to measure everything from grit and self-esteem to anxiety and burnout. These instruments shape hiring decisions, clinical diagnoses, and entire research literatures. Yet a subtle flaw has haunted survey design for decades: when questionnaires mix positively and negatively worded items, respondents often produce answers that reflect the phrasing of the questions rather than the trait being measured. A new study published in Behavior Research Methods by Palmira Faraci and Giuliana Nasonte of the Psychometrics Laboratory at the University Kore of Enna puts a sophisticated statistical tool, random intercept item factor analysis, to the test, and the results suggest it can cleanly separate genuine personality signals from the noise created by question wording.</p>
<p>The problem, known in psychometrics as method variance or the wording effect, arises because scales are frequently built with a deliberate balance of positively and negatively keyed items. The idea sounds sensible: reversing the direction of some questions should discourage automatic agreement and force respondents to read carefully. But the strategy backfires in a measurable way. Items that share the same wording direction tend to correlate with one another for reasons that have nothing to do with the underlying construct. When researchers run a factor analysis on such data, the statistical machinery often obligingly splits the scale into two clusters, one of positively worded items and one of negatively worded items, rather than the single dimension the scale was designed to capture.</p>
<p>This artificial splitting, sometimes called spurious bidimensionality, has real consequences. A researcher who concludes that a questionnaire measures two distinct traits may publish misleading findings, build theories on phantom subfactors, or compute reliability estimates that are inflated or deflated for the wrong reasons. The grit scale, a wildly popular measure of perseverance and passion for long-term goals, has been a particular battleground for this debate. Since its development by Angela Duckworth and colleagues, researchers have argued endlessly about whether grit truly has two facets, perseverance of effort and consistency of interest, or whether the apparent two-factor structure is simply an artifact of the negatively worded items embedded in the short version of the scale.</p>
<p>Faraci and Nasonte tackled this question with random intercept item factor analysis, or RIIFA, a modeling approach originally introduced by Albert Maydeu-Olivares and Donna Coffman in 2006. The core insight of RIIFA is elegant: it adds a random intercept factor to the standard factor model, a latent variable on which every item in the scale loads regardless of its wording direction. This intercept factor absorbs the shared response tendency that cuts across all items, including acquiescence, the general willingness to agree with statements, and other response styles. Once that pervasive method variance is siphoned off into its own latent dimension, the remaining substantive factor can be interpreted as the true trait, purified of the contamination introduced by how the questions happen to be phrased.</p>
<p>To evaluate whether RIIFA delivers on this promise, the researchers conducted two studies using two independent samples from the United Kingdom. The first study analyzed data from 977 participants, and the second from 496. Their test case was the Short Grit Scale, known as Grit-S, an eight-item instrument that mixes positively and negatively worded questions. The choice was strategic: the Grit-S is short, widely used, and famously prone to the wording-effect problem, making it an ideal proving ground for a method intended to restore structural validity to mixed-worded scales.</p>
<p>The first study focused on dimensionality assessment, the task of determining how many latent factors actually underlie a set of items. The researchers compared traditional factor retention techniques, including parallel analysis, a classic method that compares observed eigenvalues against those from random data, with RIIFA-based counterparts that incorporate the random intercept factor. They also employed exploratory graph analysis, a newer network psychometrics approach that treats items as nodes in a network and uses community detection algorithms, borrowed originally from network science methods like the fast unfolding of communities, to identify clusters of tightly connected items. The results were striking. Traditional retention methods consistently overestimated the number of factors, suggesting the Grit-S contained two dimensions when the theoretical expectation was one. In contrast, the RIIFA-based techniques produced unidimensional solutions, and bootstrap analyses, which resample the data thousands of times to gauge stability, showed that these solutions were considerably more stable than those obtained without controlling for the wording effect.</p>
<p>The second study shifted from exploration to confirmation, using confirmatory factor analysis to formally test competing models of the scale&#8217;s structure. The researchers fit models with and without a random intercept factor and compared them on both fit and parsimony, the principle that simpler explanations should be preferred unless complexity earns its keep. The model incorporating the random intercept factor achieved the best balance between the two. Its fit statistics were excellent by conventional standards: a root mean square error of approximation of .048, with a confidence interval spanning .022 to .072, a comparative fit index of .984, a Tucker-Lewis index of .974, and a standardized root mean square residual of .027. Values close to .95 or above on the fit indices and below about .06 or .08 on the error measures are typically considered indicative of a well-fitting model, and this model cleared every bar comfortably.</p>
<p>Reliability told a similarly encouraging story. The RIIFA model yielded a hierarchical omega of .84, a coefficient derived from bifactor-style measurement models that estimates how well the total scale score reflects a single dominant common factor after accounting for the variance absorbed by subsidiary dimensions. In practical terms, this means that once the method variance was reallocated to the random intercept factor, the substantive grit factor explained enough common variance to support meaningful interpretation of total scores. The researchers interpret this as evidence that RIIFA does not merely hide the problem; it actively redistributes explained variance between the substantive factor and the method factor, mitigating the artificial bidimensionality that plagues conventional analyses and enhancing the interpretability of the latent structure.</p>
<p>The implications reach well beyond the grit scale. Mixed-worded instruments are ubiquitous across psychology, appearing in measures of self-esteem, core self-evaluations, need for cognition, perceived stress, learning burnout, and countless clinical and organizational questionnaires. For each of these, the same dilemma recurs: researchers must decide whether an emerging second factor represents a genuine substantive distinction or a methodological artifact. Historically, that judgment has been made with tools, such as standard parallel analysis and conventional confirmatory factor analysis, that have no built-in mechanism for separating content from phrasing. The findings of this study suggest that RIIFA offers a principled alternative, and the authors explicitly recommend its application in cases where wording effects threaten the validity of psychometric measurement.</p>
<p>The study also connects to a broader movement in quantitative psychology toward modeling response styles and careless responding rather than simply hoping they wash out in large samples. Recent work has documented how even a few inconsistent respondents can confound the structure of personality survey data, how acquiescence distorts exploratory factor analysis, and how attention checks and response-style detection methods carry their own complications. Faraci and Nasonte&#8217;s contribution fits squarely into this research program, demonstrating on real data that a model explicitly designed to partition substantive and method variance can outperform traditional approaches. Importantly, the authors have made their data openly available through the Open Science Framework, and all of the R and Mplus code used in the analyses is accessible as well, lowering the barrier for other researchers to adopt the technique. For a field that depends so heavily on self-report questionnaires, a validated method for ensuring that scores reflect traits rather than question phrasing is not a technical nicety. It is a foundation for the credibility of the measurements themselves, and this study provides compelling empirical grounds for adding random intercept item factor analysis to the standard psychometric toolkit.</p>
<p><strong>Subject of Research:</strong> Using random intercept item factor analysis to separate substantive variance from wording-effect method variance in mixed-worded psychological scales</p>
<p><strong>Article Title:</strong> Disentangling substantive and method variance in mixed-worded scales: An empirical application of the random intercept item factor analysis (RIIFA)</p>
<p><strong>Article References:</strong> Faraci, P., &amp; Nasonte, G. (2026). Disentangling substantive and method variance in mixed-worded scales: An empirical application of the random intercept item factor analysis (RIIFA). <em>Behavior Research Methods, 58</em>(11), Article 299. <a href="https://doi.org/10.3758/s13428-026-03163-1" rel="noopener noreferrer">https://doi.org/10.3758/s13428-026-03163-1</a></p>
<p><strong>Image Credits:</strong> AI Generated</p>
<p><strong>DOI:</strong> <a href="https://doi.org/10.3758/s13428-026-03163-1" rel="noopener noreferrer">10.3758/s13428-026-03163-1</a></p>
<p><strong>Keywords:</strong> psychometrics, random intercept item factor analysis, wording effects, method variance, Short Grit Scale, factor analysis, exploratory graph analysis, parallel analysis, dimensionality, structural validity, confirmatory factor analysis, survey methodology</p>
]]></content:encoded>
					
		
		
		<post-id xmlns="com-wordpress:feed-additions:1">215687</post-id>	</item>
	</channel>
</rss>
