Common Method Bias: It's Bad, It's Complex, It's Widespread, and It's Not Easy To Fix
Common Method Bias: It's Bad, It's Complex, It's Widespread, and It's Not Easy To Fix
Organizational Behavior
Common Method Bias: It’s Bad,
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
17
Any measure. . .reflects not only a theoretical concept of interest but also measurement error. Mea-
surement error. . .can be partitioned into random error and systematic error, such as method variance.
Method variance refers to variance attributable to the measurement method rather than the construct
of interest. . . . Each of these two components can have serious confounding influences on empirical
research and yield misleading conclusions. . . . Because measurement errors (i.e., random error and
method variance) provide potential threats to the validity of research findings, it is important to validate
measures and disentangle the distorting influences of these errors before testing theory.
—Bagozzi et al. (1991, p. 421)
INTRODUCTION
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
As indicated by the quotation above, researchers in the organizational and behavioral sciences are
concerned about the potential harmful effects of systematic measurement error on the validity
of their research findings. Indeed, since Campbell & Fiske (1959) focused attention on this issue
65 years ago, it has been an enduring topic of discussion (e.g., Cote & Buckley 1987, Doty &
Glick 1998, Evans 1985, Hulland et al. 2018, Podsakoff et al. 2003, Spector et al. 2019). However,
despite widespread recognition that method biases can have several harmful effects, our reading
of the literature and discussions with colleagues suggest that the causes and consequences of com-
mon method bias (CMB), and the remedies for dealing with it, are still not well understood, for
several reasons. First, the effects that method factors have on the observed relationships between
variables are complex. Indeed, several sources of CMB exist, and it is difficult if not impossible to
control all of them in a single study. This complexity is compounded by the fact that as the field
has moved from examining simple research designs using regression techniques to more compli-
cated multidimensional constructs ( Johnson et al. 2011) using multilevel analyses (Lai et al. 2013,
Mathieu et al. 2012), it has become more difficult to determine how to control for the effects of
CMB.
Second, there is considerable debate about the potential impact that CMB has on the rela-
tionships between constructs. Some scholars argue that CMB is an important problem that needs
to be identified and controlled (e.g., Bagozzi 2011; Burton-Jones 2009; Cote & Buckley 1988;
Podsakoff et al. 2003, 2012; Williams et al. 2010), and others claim that the potential effects of
CMB are (at best) exaggerated (Bozionelos & Simmering 2022, Fuller et al. 2016, Spector 1987,
Spector & Brannick 2010) and (at worst) may represent an urban legend that can generally be
ignored (Brannick et al. 2010, question 2; Spector 2006). This debate produces confusion about
the consequences of CMB and the necessity of identifying remedies to control them.
Third, the sheer amount of material published on CMB makes it difficult for even the most
devoted scholars to keep up with this literature. For example, more than three dozen articles
examining techniques for assessing or controlling CMB have been published since 2010 (e.g.,
Antonakis et al. 2010, Jordan & Troth 2020, Kock et al. 2021, MacKenzie & Podsakoff 2012, Yao
& Xu 2021, Zhang et al. 2022), and these articles received more than 13,000 combined citations in
2022 alone (according to a 2023 Google Scholar search). Finally, even when researchers are aware
of the potential problems that systematic CMB can produce, they may be unclear about how to
minimize their harmful effects.
Therefore, the goal of this article is to increase our understanding of CMB’s causes, conse-
quences, and potential remedies. First, we discuss why CMB is bad, highlighting the harmful
effects it can have on the estimates of construct reliability and validity and on the relationships be-
tween measures of different constructs. Second, we explore the complex nature of CMB. To better
understand this complexity and what it means for researchers interested in controlling CMB, we
discuss its sources and the conditions under which it is likely to have its biggest effects. Third,
we summarize evidence indicating that studies using designs susceptible to CMB are relatively
18 Podsakoff et al.
widespread. Indeed, the evidence suggests that the conditions in which CMB is likely to have bi-
asing effects are quite common in several disciplines. Fourth, we discuss why CMB is not easy to
fix by reviewing the strengths and limitations of various procedural and statistical remedies used
to minimize its harmful effects. Finally, we identify several avenues for future research.
This article builds on and extends our earlier reviews (MacKenzie & Podsakoff 2012; Podsakoff
et al. 2003, 2012; Podsakoff & Organ 1986) and makes several contributions to the literature.
First, we provide an updated review incorporating research on CMB reported since our earlier
article (Podsakoff et al. 2012). Second, we provide illustrations to help clarify the potential effects
that CMB has on the reliability and validity of measures, as well as the relationships between
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
constructs. Third, we examine and critique some recent claims (Bozionelos & Simmering 2022,
Cruz 2022, Fuller et al. 2016) that CMB is not a threat to research findings. Finally, we provide
recommendations to guide researchers interested in controlling the potential effects of CMB and
discuss several avenues for future research.
CONSTRUCT A
Proportion of variance in items accounted
for by systematic (measurement) error
Figure 1
Partitioning indicator variance into component parts. Consistent with conventions in the literature (e.g.,
Bollen 1989, Brown 2015), the oval represents a latent construct and the rectangles represent reflective
indicators (items) used to measure the construct. The different colors represent the different sources of
variance accounted for in each indicator. Specifically, the variance in each indicator is partitioned into
portions attributable to the construct being measured (blue), the method used to measure the construct (red),
and random measurement error (green). The dotted blue lines reflect the average variance extracted in the
items by the underlying construct. For simplicity, we omit variance that may be unique to the items.
items. This observation is consistent with the findings of several studies, which have reported
that method factors have unequal effects on different measures, regardless of whether they are
different measures of the same construct ( Johnson et al. 2011; Podsakoff et al. 2012; Rafferty
& Griffin 2004, 2006; Williams et al. 2010) or measures of different constructs (Baumgartner &
Steenkamp 2001, Cote & Buckley 1987). Indeed, indicators can be influenced by different method
factors, and the same method factor can have stronger or weaker effects on different indicators of a
given construct or on different constructs (Campbell & Fiske 1959, McDermott & Sharma 2017,
Messick 1991, Podsakoff et al. 2012, Spector et al. 2022). Finally, Figure 1 shows that random
measurement error also varies across the items used to measure the focal construct. This variation
reflects the fact that this form of error is random and depends on the amount of variance accounted
for by the latent construct and the methods used to measure it.
Figure 2
Example of systematic measurement error biasing estimates of construct validity and reliability. The blue
circles represent systematic variance in the items that is attributable to the latent construct and any method
characteristics shared by the items.
20 Podsakoff et al.
any method characteristics shared by the items. When variance attributable to method factors is
not controlled and is lumped together with construct variance, it can bias the estimates of con-
struct validity and reliability. The biasing effects of systematic measurement error due to common
methods on estimates of construct validity and reliability produce several potential problems.
First, method factors that inflate or attenuate interitem covariation will bias estimates of factor
loadings, reliability coefficients, and AVE estimates (MacKenzie & Podsakoff 2012, Podsakoff
et al. 2012), which can lead to incorrect conclusions about the adequacy of a scale’s reliability and
item-level convergent validity (Bagozzi 1984; Baumgartner & Steenkamp 2001; Brannick et al.
2010; Cote & Buckley 1987; Podsakoff et al. 2003, 2012; Williams et al. 1989, 2010). Indeed,
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
when systematic method variance in the items is not controlled, researchers may overestimate
the reliability and validity of their scales and, in more extreme cases, conclude that their scales
are reliable and valid when in fact they are not. Second, common method variance (CMV) can
produce inaccurate “corrected” correlations in meta-analyses when the reliability estimates used
to calculate the correction are biased (Le et al. 2009, MacKenzie & Podsakoff 2012, Podsakoff
et al. 2012). Specifically, reliability-corrected correlations will understate the relationship between
focal variables when reliability estimates are inflated by CMB, and these estimates will overstate
the relationship between these variables when reliability estimates are attenuated by CMB.
Table 1 Summary of studies reporting CFAs of MTMM matrices to partition trait, method, and error variance in latent variables
Variance Variance Variance
22
attributable to trait attributable to attributable to
Reference Sample Estimation technique (construct) factors method factors random error
Cote & Buckley 70 matrices examining a wide variety of Traditional CFA of MTMM 42% 26% 32%
(1987) constructs from the fields of marketing, matrices (CFA-MTMM)
psychology/sociology, education, and other
Podsakoff et al.
business disciplines
Williams et al. 11 matrices involving perceptions of jobs and Traditional CFA of MTMM 48% 25% 21%
(1989)a work environments matrices (CFA-MTMM)
Buckley et al. (1990) 61 matrices examining a variety of constructs Traditional CFA of MTMM 42% 22% 36%
matrices (CFA-MTMM)
Doty & Glick (1998) 28 matrices examining constructs from a Traditional CFA of MTMM 46% 32% 22%
variety of social science disciplines matrices (CFA-MTMM)
Mishra (2000) 6 matrices examining health care–related Traditional CFA of MTMM 36% 30% 34%
constructs matrices (CFA-MTMM)
Ketokivi & Schroeder Data from 164 manufacturing plants from five CFA of MTMM matrices 36% 18%b 46%
(2004) countries (Germany, Italy, Japan, United using a CTUM
Kingdom, United States) in three industries
(automotive supply, machinery, electronics)
on four measures of market performance
(price, product quality, product image,
product features) taken from three sources
(plant manager, plant superintendent, plant
research coordinator)
Lance et al. (2010) 18 matrices CFA of MTMM matrices 40% 18% 42%
using a CTCM
Averages 41% 24% 33%
a
Values reported for variance estimates represent medians.
b
The average amount of method variance per source varied depending on the performance measure used (e.g., average method variance for price = 4%; quality = 25%; product image = 20%;
product features = 23%), providing additional evidence that the amount of method variance differs across constructs.
Table adapted with permission from Podsakoff et al. (2012).
Abbreviations: CFA, confirmatory factor analysis; CTCM, correlated traits–correlated methods model; CTUM, correlated traits–uncorrelated methods model; MTMM, multitrait multimethod.
SUPPORTIVE
LEADER
BEHAVIOR
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
EMPLOYEE
HELPING
BEHAVIOR
Figure 3
Example of systematic measurement error biasing the relationship between constructs. Each of the two
correlated latent variables (supportive leader behavior and employee helping behavior) is measured by three
indicators using self-reports obtained from the same source (employees).
of two constructs when the correlation between the method factors is lower than the observed
correlation between the measures with method effects removed, (b) inflate the relationship when
the correlation between the method factors is higher than the observed correlation between the
measures with method effects removed, and (c) have no influence on the relationship when the
correlation between the method factors is the same as the observed correlation between the vari-
ables with the method effects removed. However, regardless of whether the method factor inflates
or deflates the relationship, it can cause several serious problems (Bagozzi 1984; Baumgartner &
Steenkamp 2001; MacKenzie & Podsakoff 2012; Podsakoff et al. 2003, 2012; Siemsen et al. 2010).
First, since systematic measurement error can inflate or deflate the estimates of the observed
relationships between two latent variables, these errors may lead a researcher to conclude either
that a relationship between the two variables exists when it does not (Type I error) or that a rela-
tionship does not exist when indeed it does (Type II error). Second, if the predictor and criterion
variables share systematic error variance, the amount of variance accounted for in the criterion
variable(s) by the predictor variable(s) may be either understated or overstated. Finally, because
systematic measurement error can inflate or deflate the estimates of the observed relationships
between latent variables, it can enhance or attenuate the observed relationships between a fo-
cal construct and its antecedents, correlates, and consequences, and subsequently influence the
are obtained from the same (as opposed to different) raters, or the potential effects that research
designs (cross-sectional versus lagged designs) have on these relationships. Still other evidence
comes from research on the effects that item characteristics, item contexts, or measurement con-
texts have on CMB. As a starting point, we compare the differences in the correlations between
variables when they were obtained from the same (versus different) sources and at the same (versus
different) times (we return to the effects of other factors in the section titled The Effects of Com-
mon Method Bias Are Complex). To obtain estimates of the effects of CMB on the covariation
between constructs, we analyzed correlations reported in published meta-analyses. We searched
for and coded meta-analyses that reported correlations between measures of constructs rated by
the same and different sources or rated at the same and different times. Our analyses included
233 bivariate correlations from 59 meta-analyses for the effects of rating sources as well as
236 bivariate correlations from 33 meta-analyses for the effects of rating times (for details on
the literature search, inclusion criteria, and coding and the list of meta-analyses included in our
analyses, see Supplemental Appendix A).
Although several researchers (e.g., Baumgartner et al. 2021, Podsakoff et al. 2003) have ob-
served heterogeneity in the effects of CMB, few studies have attempted to identify the factors that
predict this variability. To help determine whether correlations of various relationships are subject
to CMB to the same extent, we coded the valence (positive versus negative) of the predictor and
criterion variables of each relationship in our data set and examined how it influenced the observed
correlations. Earlier research (Kam & Meyer 2015, Magazine et al. 1996, Zeng et al. 2020) showed
that positively and negatively worded items influence both construct dimensionality and nomo-
logical validity. Examples of predictor variables categorized as having a positive valence include
positive individual differences [e.g., conscientiousness, agreeableness, core self-evaluation (CSE)],
work attitudes (e.g., job satisfaction, organizational commitment), task characteristics (e.g., job
challenge, autonomy), and leadership behaviors (e.g., transformational leadership, servant leader-
ship). In contrast, examples of predictor variables categorized as having a negative valence include
negative individual differences (e.g., neuroticism, negative affect), work attitudes (e.g., burnout,
cynicism about change), leadership behaviors (e.g., abusive supervision, unethical leadership), and
job stressors or strains. Positive criterion variables include positive attitudes (e.g., commitment,
satisfaction, engagement) or behaviors/performance [e.g., task performance, organizational citi-
zenship behavior (OCB), creativity]. Negative criterion variables include negative attitudes (e.g.,
negative affect), perceptions (e.g., role ambiguity, role conflict), and behaviors [e.g., counterpro-
ductive work behavior (CWB)]. For a complete list of variables included in each valence category,
see Supplemental Appendix B.
After coding the valence of the predictor and criterion variables, we examined four categories
of relationship valence: positive–positive, positive–negative, negative–negative, and negative–
positive. For each of these categories, we obtained the weighted average inflation rates in the
correlations of same-source ratings compared with different-source ratings as well as in the cor-
relations of cross-sectional designs compared with lagged designs (i.e., ρ same source /ρ different source
24 Podsakoff et al.
Percentage of inflation
Same versus different sources
Same versus different times
400
350
198%
300 (78%, 281%) 251%
Percentage of inflation
(101%, 348%)
165%
250
(109%, 218%)
161%
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
50
0
Positive–positive Positive–negative Negative–negative Negative–positive
Valence of relationships
Figure 4
Effects of rating source and temporal separation on correlations across relationships of different valence. The number above each bar is
the average inflation rate, followed by its twentieth-to-eightieth-percentile range. The blue bars compare the weighted average
correlations of same-source versus different-source ratings across valence categories. The red bars compare the weighted average
correlations of same-time versus different-time ratings across valence categories. For details on the meta-analytic data used in these
estimates, see Supplemental Appendix C, Table C1 (same versus different sources) and Table C2 (same versus different times).
or ρ same time /ρ different time ), weighted by the number of studies included in the meta-analyses
we coded.
26 Podsakoff et al.
Rater characteristics
Figure 5
Illustration of the primary sources of common method bias.
presence of implicit theories), or are given the opportunity (or have the motivation) to satisfice
in responding. On the other hand, item characteristics and item context effects are likely to be
problems when respondents lack motivation (because the repetitiveness of items or the scale is
too long), the questionnaire is too difficult (e.g., questions are complex, abstract, or ambiguous or
require retrospective recall), or respondents are given the opportunity or the motivation to sat-
isfice (e.g., because scales contain common properties, items measuring the same constructs are
grouped together, or answers to similar questions are in close proximity).
Finally, measurement context is a concern when respondents lack the motivation to respond
(e.g., because the context arouses suspicions about the researcher’s intent, the source of the survey
is disliked or distrusted) or because the task is too difficult (e.g., the survey is conducted over the
phone and does not allow the respondent to read the questions). Therefore, researchers interested
in controlling for CMB need to pay attention not only to the various sources of CMB but also to
factors such as the respondents’ ability and motivation and the difficulty of the survey, which are
likely to heighten the effects of these biases when they are present.
Leviatan 1975, Lord et al. 1978). Implicit theories have also been
used to help explain the relationships between job satisfaction
and job performance (Smither et al. 1989), attributions of the
causes of group performance (Bachrach et al. 2001, Staw 1975),
employee silence (Detert & Edmondson 2011), and OCBs and
performance evaluations (Podsakoff et al. 2013).
Consistency motif Tendency for respondents to try to Respondents asked to rate items that reflect similar content areas
maintain consistency in their (either within or between constructs) will try to maintain
responses to similar items on a consistency in their ratings, thereby inflating the covariation
questionnaire or to organize between these items or constructs (Podsakoff & Organ 1986,
their responses in a consistent Schmitt 1994). These biases may be particularly likely to occur
manner when respondents are asked to provide retrospective accounts of
their attitudes, behaviors, and perceptions (Podsakoff et al. 2003).
Social desirability Tendency for respondents to The correlation between one variable and another may be
respond to items in a way that influenced by the respondents’ desire to be viewed in a positive
puts them in a favorable light or light. Among the topics that have typically been shown to be
that is viewed favorably by influenced by social desirability are self-reports of personality
others, rather than on the basis traits (Edwards 1957) and feelings of self-esteem (Astra & Singg
of their real feelings 2000, Greenwald & Farnham 2000, Huang 2013). However,
social desirability has generally not been shown to be related to
job performance (Ones et al. 1996).
Leniency biases Tendency for respondents to Respondents are likely to rate the personality, attitudes, beliefs, and
provide ratings of themselves or intentions of people they like more favorably than those of
others that are more favorable people that they dislike. In addition, Cheng et al. (2017) report
than is warranted that a rater’s personality is related to leniency biases; they found
that agreeableness and extroversion are positively related to
rating other people leniently and that agreeableness and
conscientiousness are positively related to self-ratings, whereas
neuroticism is negatively related to self-ratings.
Response styles Tendency for respondents to Baumgartner & Steenkamp (2001) found that five different
systematically differ in the use of response styles (acquiescent, disacquiescent, extreme, midpoint,
response scales and noncontingent response styles) accounted for 27% of the
Common response styles include variance in the magnitude of the correlations between 14
the tendency to acquiesce (agree) consumer constructs. They also found that the correlation
or to disacquiesce (disagree) with between constructs can be biased upward or downward
items, irrespective of the content depending on the correlation between the response styles.
of the item, or to use extreme, Moreover, Weijters et al. (2010) found evidence that acquiescent
midpoint, or noncontingent scale and extreme response styles are largely consistent over the course
points of a survey. To the extent that these tendencies account for
systematic variance in the relationships between variables that is
different from the true score (actual) variance that exists between
these variables, it is problematic.
(Continued)
28 Podsakoff et al.
Table 2 (Continued)
Potential cause Definition Examples/evidence of effects
Positive/negative Tendency for respondents to view Several researchers (Connolly & Viswesvaran 2000, Thoresen et al.
affectivity, themselves or the world around 2003) have noted that the covariation between items (or
emotionality, or mood them in positive emotional terms constructs) on a questionnaire may be influenced by respondents’
(e.g., enthusiasm, joy, energy, tendencies to view the world in generally positive (or negative)
cheerfulness) or in negative terms, irrespective of the content of the items. To the extent that
emotional terms (e.g., fatigue, these tendencies account for systematic variance in the
sadness, disgust, distress) relationships between variables that is different from the true
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
(Continued)
agreement, frequency, similarity) necessary to process the information contained in the question
and be more likely to exhibit undifferentiated responses,
subsequently increasing the consistency of responses across the
survey items and the likelihood of method biases. However,
neither these authors nor Spector & Nixon (2019) found
evidence that common scale formats produce stronger
relationships between constructs than different scale formats.
Common scale anchors Items on a questionnaire written Podsakoff et al. (2013) argue that repeated exposure to the same
with the same scale anchors (e.g., scale anchors decreases respondents’ motivation to exert the
“strongly disagree” to “strongly cognitive effort necessary to process the information contained
agree”) or the same number of in scale items and increases the probability of undifferentiated
anchor points responses, which subsequently increases the consistency across
scale items and the likelihood of method biases. Consistent with
this explanation, these authors reported that estimates of the
relationship between OCBs and performance evaluations were
39% larger when studies used the same (versus a different)
number of anchor points when assessing both constructs.
Positively and negatively Items on a questionnaire written Several studies (Greenberger et al. 2003, Harvey et al. 1985,
worded items with the same evaluative Ibrahim 2001) have demonstrated that negatively worded items
(positive or negative) wording often produce method factors that are composed solely of
negatively worded items, raising concerns about the construct
validity of the measures. Schmitt & Stults (1986) note that these
effects may result from the fact that once respondents establish a
pattern of responding to survey items, they may ignore the
positive–negative wording of the items.
Item context effects Covariation observed between items (or variables) produced by their relationship to one another on a
questionnaire
Item priming effects The placement of items Janiszewski & Wyer (2014) have noted that the exposure to an
(constructs) on a questionnaire initially encountered stimulus (item) makes the processing of that
may make subsequent items stimulus more accessible and may subsequently influence all
(constructs) on the questionnaire stages of the survey response process, including attention,
more salient and imply a comprehension, memory retrieval, inference, and response
relationship between the generation. Accordingly, researchers ( Judd et al. 1991,
variables. Tourangeau et al. 1991) have reported that prior questionnaire
items affect the speed with which respondents answer similar
subsequent questions and that the size of this effect is a function
of how closely related the subsequent items are to the initial item
prime.
(Continued)
30 Podsakoff et al.
Table 2 (Continued)
Potential cause Definition Examples/evidence of effects
Item embeddedness Items embedded among other Harrison & McLaughlin (1993) reported that evaluatively neutral
effects positively or negatively worded items that were positioned in blocks of positive (or negative)
items will take on the evaluative evaluative items were rated similarly to the items they were
properties of those items. embedded in. In a follow-up study, Harrison et al. (1996)
reported that the correlation between subjects’ ratings of
outcome favorability and perceptions of fairness was only 0.10 in
a positive measurement context but increased to 0.50 in a
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
(Continued)
memory, and (c) facilitate the use of implicit theories when they
exist.” This observation suggests that relationships among items
(or constructs) measured at the same point in time will be
stronger than when measured at different points in time.
Predictor and criterion Items measuring the predictor and Responses obtained in the same location reduce the likelihood of
variables measured in criterion variables may be differential cues in the environment that might cause distractions
the same location obtained from respondents in for respondents, thereby strengthening the relationships between
the same location. the measures.
Predictor and criterion Items measuring the predictor and Responses obtained in the same medium reduce the likelihood of
variables measured criterion variables may be differential cues that might cause distractions for respondents,
using the same obtained from respondents using thereby strengthening the relationships between the measures.
medium the same medium (paper and
pencil, computer screen,
interview, etc.).
Abbreviations: MMPI, Minnesota Multiphasic Personality Inventory; OCB, organizational citizenship behavior.
( Jakobsen & Jensen 2015, Meier & O’Toole 2013); tourism, hospitality, supply chain, and sports
management (Kaltsonoudi et al. 2021, Kaufmann & Saw 2014, Kock et al. 2021, Min et al. 2016,
Montabon et al. 2018, Zhu et al. 2022); marketing (Hulland et al. 2018); and management in-
formation systems (Cram et al. 2019) (for a summary, see Supplemental Appendix D). These
studies examined almost 13,000 articles and found that between 31% and 98% (and, on average,
almost 70%) of the studies reported in the articles published across these disciplines are poten-
tially susceptible to the effects of CMB, either because they obtained the focal variables from the
same source at the same point in time or because the studies were susceptible to one or more other
sources of CMB. Thus, it appears that researchers in a variety of disciplines have good reason to
be concerned about the potential effects of CMB.
32 Podsakoff et al.
Table 3 Summary of the potential causes of common method bias and circumstances under which it is likely to be a
problem
Lack of ability,
Causes education, or Lack of motivation to Task (questionnaire) is Opportunity or
associated with experience respond accurately too difficult motivation to satisfice
Rating source Lack of verbal Low personal relevance of Not applicable The availability of answers
ability, education, the issue to previous questions (in
or cognitive Low self-efficacy to provide memory)
sophistication a correct answer
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Table adapted with permission from MacKenzie & Podsakoff (2012) and Podsakoff et al. (2012).
cited (Bozionelos & Simmering 2022, Cruz 2022, Fuller et al. 2016) as support for claims that
CMB is generally not a problem in the organizational sciences.
What Are the Most Common Remedies for Dealing with Method Bias?
Almost a dozen studies have reported on the statistical and procedural techniques used by re-
searchers to control for CMB (see Supplemental Appendix E). Although these studies vary
considerably in the amount of detail they provide, several points are worth noting. First, it is
obvious that some academic fields devote more attention to, and are more aware of, the potential
harmful effects of CMB than others. For example, whereas between 59% and 81% of articles in
marker variable (MV) technique (e.g., a correlation-, regression-, or CFA-based marker) and the
unmeasured latent variable (UMLV) technique. Few studies use the directly measured latent vari-
able (DMLV) technique or the instrumental variable technique or attempt to control for response
styles. Third, although procedural remedies tend to be used somewhat less often than statisti-
cal remedies in most disciplines (for an exception, see Bozionelos & Simmering 2022), obtaining
measures of the focal variables from different sources and using temporal separation are the most
common procedural techniques, followed by proximal or psychological separation or attempts to
mitigate biases associated with similar item characteristics (e.g., same versus different scale type,
scale anchor points). Finally, Bozionelos & Simmering (2022) reported that almost 7% of the
studies they examined in the human resources domain used an unknown test and another 3.5%
simply examined the correlation matrix to determine the impact of CMB. Thus, more than 10%
of the studies in this domain did not provide adequate information about the potential effects of
CMB.
Obtaining measures from different sources. Obtaining measures of the predictor and criterion
variables from different sources (e.g., other people, objective measures, archival data) breaks the
connection between the measures of these constructs that were observed to be “correlated” due
to common rater effects (Figure 6a). The objectives of this technique are (a) to decrease biases
associated with implicit theories, consistency motifs, and transient mood states; (b) to reduce
tendencies to respond in a socially desirable or lenient manner across the measures of the focal
constructs; and (c) to minimize the effects of gathering the data at the same time in the same
location using the same medium (Podsakoff et al. 2003, 2012). Evidence of the effectiveness of
this remedy is depicted in Figure 4, which indicates that gathering measures of focal variables
from a different (as opposed to the same) source reduces the correlation (on average) between
these variables to between 160% and 250%. Thus, when appropriate, obtaining measures of the
focal variables from different sources is an effective remedy to the effects of CMB.
Traditionally, alternatives to self-reports include objective indicators (e.g., number of units pro-
duced, sales dollars, percentage of sales quota, number of days absent, voluntary turnover) and
ratings from other individuals (e.g., supervisors, peers, customers, or on occasion a spouse or a
34 Podsakoff et al.
a Obtain measures of predictor and b Separate the measurement of the predictor and criterion
criterion variables from different sources variables temporally, proximally, or psychologically
CONSTRUCT A CONSTRUCT B
CONSTRUCT A CONSTRUCT B
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Figure 6
Illustrations of the basic types of procedural remedies. (a) Obtaining measures of predictor and criterion variables from different
sources. (b) Separating the measures of the predictor and criterion variables temporally, proximally, or psychologically. (c) Protecting
respondent anonymity and reducing evaluation apprehension. (d) Minimizing common scale properties. The dotted red lines represent
the broken connections between measures of two constructs that were observed to be “correlated” due to common rater effects.
significant other). For example, in their study of the moderating effect of long-term orientation on
the relationship between corporate social responsibility (CSR) and new venture financial perfor-
mance, Wang & Bansal (2012) measured long-term orientation and financial performance from
CEOs and presidents of new ventures and CSR from the firms’ websites. As another example,
Boehm et al. (2014) gathered measures from six different sources to examine the effects of age-
inclusive human resource practices on the development of an organization-wide diversity climate,
collective perceptions of social exchange, firm performance, and employees’ collective turnover
intentions. However, researchers have begun to use techniques employed in other disciplines to
go beyond the use of self-report measures. For example, Ganster & Rosen (2013) reviewed the
use of physiological measures from the field of medicine (e.g., blood pressure, cholesterol, and
body mass index) to assess employee well-being, and Chaffin et al. (2017) explored the utility of
wearable sensors to obtain measures of behavioral constructs at both the individual and group lev-
els. Although Chaffin et al. caution researchers to be aware of the potential sources of error when
using these sensors, they also highlight the possible benefits of working together with the devel-
opers of these devices for the purpose of extending our current toolkit. We agree and encourage
tive data such as job performance is difficult because some jobs (e.g., white collar and managerial
jobs) may not have readily available objective measures (Hopp et al. 2009), objective indicators
may capture only some of the important aspects of overall job performance (Rotundo & Sackett
2002), and organizations may be unwilling to share some forms of archival data (e.g., personnel
records) because of legal or other constraints. Finally, obtaining data from different sources and
matching the information from the predictor and criterion variables generally take more time,
energy, and resources than does obtaining data from only one source.
A variation of this remedy that is often used in research on groups and teams is to randomly
split a sample of employees from a single work unit into two groups and then to gather the predic-
tor variable(s) from one of the subsamples and the criterion variable(s) from the other subsample.
This split-sample design has been operationalized in different ways. Some studies (e.g., Smith et al.
1983, Zohar & Luria 2004) split their overall sample to obtain aggregated measures of the predic-
tor variables from the first subsample and obtain aggregated measures of the criterion variables
from the second subsample. In doing so, they lessened the impact of same-source biases on the
estimates of the substantive relationships between predictors and criteria. Alternatively, the split-
sample technique can be applied to the measurement of multidimensional constructs to remove
same-source biases from the ratings of different dimensions of the same construct. For example,
Schulte et al. (2009) split participants of each unit into seven groups and collected ratings of one
of seven different climate subdimensions from each group to alleviate concerns that response bias
could inflate the interdimensional relationships.
Although splitting the sample using either of these techniques has the advantage of gather-
ing the predictor and criterion variables from different sources, these remedies are not without
limitations. First, as noted by several authors (Bliese 2000, Morgeson & Hofmann 1999), re-
searchers should not assume without good theoretical reasons that the structure or function of
individual-level constructs generalizes to the group or organizational level (i.e., that the con-
structs are isomorphic). Second, it is important for the referent of the measures to reflect the
level of the construct (individual or group) being operationalized (Chan 1998, Podsakoff et al.
2015). Third, with some composition models, the aggregation of the focal variables must be jus-
tified statistically by demonstrating that a substantial, meaningful amount of variance in measures
of the focal variables is attributable to between-group factors (Bliese 2000, Chan 1998). Fourth,
aggregating both the predictor and criterion variables will reduce the sample size and the sub-
sequent power of the statistical test used, thereby increasing the likelihood of Type II errors
(Podsakoff & Organ 1986). Finally, using the responses of coworkers as surrogates for individual-
level ratings of some constructs is questionable. For example, aggregating coworker responses of
their leaders’ behavior is inconsistent with research showing that leaders respond differently to
different employees (Henderson et al. 2009, Martin et al. 2018). However, assuming that there
are no theoretical or methodological restrictions from aggregating the data, these remedies do
have advantages for those concerned with CMB associated with gathering data from the same
source.
36 Podsakoff et al.
Introducing a separation between the predictor and criterion variables. When it is impos-
sible or inappropriate to obtain measures of the focal variables from different sources, another
procedural remedy that aims to reduce CMB is to introduce a separation between the predic-
tor and criterion variables included in the study (Figure 6b). Although this technique acquires
measures of the focal variables from the same source, it separates these measures temporally,
psychologically, or proximally. This remedy may prove particularly useful when examining the re-
lationships between internal states (e.g., attitudes, beliefs, moods, values, perceptions, intentions)
or behaviors that are difficult for other individuals to observe because they have a low base rate or
are purposely hidden from view (e.g., sabotage, retaliation, unethical behavior).
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Our review of meta-analytic studies that examined the effects of temporal separation between
the predictor and criterion variables (Figure 4) indicates that, except for positive–negative valence
pairs, gathering measures of these variables at different points (versus the same point) in time de-
creases the correlation (on average) between 121% and 142%. These findings suggest that another
way of reducing CMB is to introduce a temporal separation between the focal variables. Examples
of the use of temporal separation include studies conducted by Gielnik et al. (2018), who exam-
ined the moderating role of age-related factors (future time perspective and prior entrepreneurial
experience) on the relationships among entrepreneurial identification, entrepreneurial intentions,
and self-reported entrepreneurial activity in a three-wave study, and by Loi et al. (2020, study 1),
who examined the sequential mediating effects of moral licensing (moral credits and moral cre-
dentials) and psychological entitlement on the relationships between employee volunteering and
workplace deviance in a four-wave study.
However, this remedy has limitations. First, since we lack a comprehensive understanding
about the length of time over which predictor variables have their effects in our field (Mitchell
& James 2001, Shipp & Cole 2015, Shipp & Jansen 2021), the lags we use in our studies may be
either too short (thus proving ineffective) or too long (thereby allowing other factors to influence
the outcome variables or the focal effect to dissipate) (Podsakoff et al. 2012, Spector 2019). This
lack of understanding is compounded by the fact that different phenomena are likely to require
different temporal delays. Second, in our role as reviewers, we have encountered studies that em-
ploy time delays as short as 1 h. This short duration is problematic, because research ( Johnson
et al. 2011, Ostroff et al. 2002) has shown that although 3-week to 1-month delays can cause an
appreciable reduction in the relationships between focal variables, a 1-h delay does not. Third,
implementing a temporal delay assumes that the relationships between the predictor and crite-
rion variables are stable over time; otherwise, conclusions about the dissipating effects of CMB
may be incorrect. This may be particularly important in the case of some variables, such as em-
ployee affect (Beal et al. 2005), moods (Scott et al. 2020), and emotions (Lim et al. 2018), which
fluctuate considerably over time (Podsakoff et al. 2019). Moreover, gathering data over time may
result in respondent attrition and is likely to take additional time and energy on the part of the
research team. Finally, although this remedy should help minimize the potential effects of some of
the sources of bias related to raters, including consistency motifs and transient mood states, as well
as some of the item-related sources of bias, it is unclear whether it counters the effects of leniency
biases, response styles, or trait social desirability. That may be why our analysis (Figure 4) found
that temporal separation had less of an effect on the relationships between focal variables than did
the use of measures from different sources. Despite these limitations, temporal separation is likely
not only to enhance the causal inferences between the predictor and criterion variables but also
to assuage concerns about the effects of CMB.
Of course, a particularly effective way of controlling for CMB is to obtain measures of the
focal variables from different sources and include a temporal separation between them. For
example, Sessions et al. (2020, study 1) examined the positive and negative effects of employee
38 Podsakoff et al.
the preliminary evidence supporting the efficacy of this procedure is encouraging, and we believe
it would prove very worthwhile for researchers to conduct additional studies that compare this
technique with obtaining measures from different sources and different time waves to determine
its generalizability.
Weijters et al. (2009) present evidence of the effects of proximal separation. These authors
examined the effects of item proximity and the nature of the conceptual relationship (unre-
lated, reversed, or nonreversed conceptual meaning) between survey items on the strength of
item correlations. They found that the correlation between item pairs measuring unrelated con-
structs increased by 225% (from 0.04 to 0.09) when these items were positioned next to one
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
another compared with when they were positioned six items apart. In addition, they reported that
(a) the average correlation between nonreversed item pairs that were positioned six or more items
apart increased by 177% (from 0.35 to 0.62) when they were positioned next to one another, and
(b) the average correlation between reversed item pairs that were positioned six or more items
apart decreased by 433% (from −0.26 to −0.06) when they were positioned next to one another.
The latter findings suggest that the positive correlation between item pairs becomes weaker for
nonreversed items, while the negative correlation for reversed item pairs becomes stronger, the
further the items are positioned from one another (Weijters et al. 2009).
Johnson et al. (2011, study 2) also examined the effects of proximal separation by adding filler
scales between the measures of their higher-order CSE construct. They found that doing so re-
duced loadings on the higher-order CSE construct by an average of 12% and decreased the R2
for the relationships between CSE and job satisfaction by 31%. However, in addition to adding
filler scales, these authors changed the response formats of the scales; therefore, it is not possible
to isolate the unique effects of proximal separation in their study. Nevertheless, taken together,
the studies by Weijters et al. (2009) and Johnson et al. (2011) suggest that researchers who are
interested in reducing the effects of CMB may want to consider either dispersing indicators of the
same construct throughout the questionnaire, rather than grouping the items together, or adding
buffer items between measures of the same construct. Of course, inserting buffer items may reduce
the motivation of respondents and increase their fatigue, and using this remedy is likely to reduce
factor loadings and reliabilities of the focal constructs. Nevertheless, Weijters et al. (2009) have
shown that inserting even a few spaces between items has a nontrivial effect on reducing CMB.
Ensuring respondent anonymity. Another technique that has been used to reduce the effects of
CMB is to protect respondents’ anonymity (Figure 6c). The purpose of this technique is to reduce
the evaluation apprehension that respondents might experience in providing their responses by
not asking for personal information that identifies them. Although guaranteeing anonymity should
lessen respondents’ motivation to edit their responses to be more socially desirable, lenient, and
consistent, this procedure is not likely to remove biases based on implicit theories, item charac-
teristics, or item context effects. Moreover, if anonymity is provided for only one, but not both,
of the focal variables (i.e., either the predictor or the criterion variable), it makes it difficult if not
impossible to match the responses obtained using this technique with data obtained from other
sources or over time, unless a linking variable that is not associated with the respondent’s identity
is used (Podsakoff et al. 2003). Finally, research by Lelkes et al. (2012) suggests that although com-
plete anonymity may reduce a respondent’s motivation to distort responses in a socially desirable
direction and subsequently increase reports of socially undesirable behavior, it may also decrease
accountability, reduce reporting accuracy, and increase satisficing. Therefore, researchers should
not routinely assume that providing anonymity to respondents reduces CMB.
Minimizing common scale formats/properties and reducing item social desirability. The
fact that several sources of CMB originate from item characteristics and item context effects
when studies used the same number (versus a different number) of anchor points. These findings
suggest that varying the number of anchor points across measures included in a questionnaire may
help reduce CMB associated with item similarity.
The effects of scale formats are not as straightforward. For example, Podsakoff et al. (2013)
did not find support for the hypothesis that common scale formats produce stronger relation-
ships than different scale formats, and Spector & Nixon (2019) reported only mixed support for
the expectation that scale formats produce differences in the correlations between stress-related
constructs. In contrast, both Dalal (2005) and Spector et al. (2010) reported that OCB and CWB
are more strongly related when agreement scales rather than frequency scales are used. Thus, ad-
ditional research on the effects of this methodological procedure is warranted. In any case, note
that although the use of some scale formats (e.g., frequency scales) may be appropriate for some
constructs (e.g., behaviors), they may be less appropriate for other constructs (e.g., attitudes); thus,
researchers also need to ensure that their scale formats match the nature of the focal constructs
they are measuring.
Reverse-coding items is designed to minimize acquiescence bias and reduce respondents’ ten-
dency to satisfice and respond stylistically by introducing “cognitive speed bumps” that require
extra processing time by the respondents. However, there is evidence that this technique may pro-
duce spurious factor structures and reduce estimates of construct reliability (Chyung et al. 2018,
Weijters & Baumgartner 2012), suggesting that this remedy should be used with caution. On a
positive note, Mathews & Shepherd (2002) reported that forewarning respondents about the pres-
ence of negatively worded items in the questionnaire reduced the amount of careless responding
and the amount of negative factor loadings in their study—but did not eliminate them. In addi-
tion, Chyung et al. (2018) noted that grouping together negative and positive items from the same
construct on a questionnaire should focus respondents’ attention and increase the probability that
respondents will process the items more deeply. Nevertheless, researchers should be aware of the
potential advantages and disadvantages of introducing negatively worded items into their survey
questionnaires.
Finally, several studies (Chen et al. 1997, Cui et al. 2022, Thomas & Kilmann 1975) have shown
that judges’ ratings of item social desirability are strongly related to the endorsement of these
items by survey respondents. In addition, Cui et al. (2022) found that (a) self- and peer ratings of
personality were equally susceptible to item social desirability and (b) the effects of item social de-
sirability were more pronounced when respondents scored high on trait social desirability. These
findings suggest that it is important to minimize item social desirability, where possible. This may
be accomplished by using neutral (as opposed to socially desirable) items in the questionnaire
(Nederhof 1985) or by removing items that correlate highly with scores on a trait social desirabil-
ity scale (Kam 2013). Note, however, that many constructs in applied psychology and management
have positive (job satisfaction, work engagement, helping behavior) or negative (neuroticism, de-
viant behavior) connotations, and it may be difficult if not impossible to develop valid measures
of these constructs while also completely eliminating item social desirability. Moreover, although
40 Podsakoff et al.
controlling for item social desirability is possible when developing and validating a scale, this
technique is less appropriate when using scales that have already been reported in the literature,
because doing so may compromise the validity of these measures. Therefore, although trying to
minimize item social desirability in the early stages of scale development is worthwhile, researchers
need to temper their desire to control this source of bias if it means compromising the validity of
their measures.
The procedural remedies discussed above attempt to lessen the effects of CMB that stems from
rating sources, characteristics of the items or the context in which the items are presented, and/or
the measurement context. Although these remedies are valuable for addressing CMB, the fact
that such biases are more likely to be present when the ability or motivation of the respondent
is lacking, the task (i.e., the questionnaire) is difficult, or the opportunity to satisfice is available
(Krosnick 1991, 1999; MacKenzie & Podsakoff 2012; Podsakoff et al. 2012) suggests that pro-
cedural remedies focused specifically on these conditions should prove valuable to researchers
interested in minimizing the effects of CMB. Table 4 summarizes the remedies that are tied to
these conditions.
Remedies directed at a lack of ability. When researchers are concerned that respondents may
lack the requisite cognitive ability or experience in dealing with the topic of interest, they should
(a) select respondents with the necessary ability, education, and familiarity with the focal topic;
(b) pretest the questionnaire with participants from the same subject pool; and (c) use terminology
and grammar that match the respondents’ capabilities (Table 4). The key here is to ensure that
there is a match between the abilities and experiences of the respondents and the questions being
asked.
Lack of motivation to Choose a sample for which the topic studied is relevant, interesting, and/or
respond accurately salient.
because respondents do Enhance the motivation to answer accurately by explaining how answers to
not understand the the questions have important consequences for the respondent and/or the
importance of the organization.
information Explain how much others (or the organization) are depending on the accuracy
of their responses.
Emphasize to respondents that their personal experiences are important, and
that only they can provide them to the researcher.
Enhance the motivation for self-expression by explaining that “we value your
feedback,” “your opinion is important to us,” and/or “we want to know what
you think.”
Treat participants in a respectful manner, let them know you value their time,
and express your appreciation for their participation.
Promise to provide feedback to respondents to motivate them to respond
accurately so that they can gain self-awareness and self-understanding.
Personalize the appeal to complete the questionnaire by adding a signature
from the researcher.
Lack of motivation to To alleviate suspicions about who will have access to the data, explain why the
respond accurately information is being gathered and how it will be used, that it is for research
because of suspicions purposes only, that only aggregated information will be provided to the
about the measurement organization, and that no individual responses will be given to anyone
context associated with the organization.
Where possible, use anonymity.
Lack of motivation to Minimize the repetitiveness of items and their grammatical redundancy.
respond accurately To the extent possible, minimize the length of the questionnaire (without
because of item compromising the construct validity of the measures).
characteristics or item
context
Task difficulty Questionnaire is too Use clear, concise language.
complex, ambiguous, or Simplify complex or abstract concepts and questions.
difficult Avoid double-barreled questions.
Clarify any vague concepts with examples.
Simplify questions and response options.
Opportunity to Opportunity to exhibit less Where possible, vary scale types, scale points, and anchor points.
satisfice cognitive effort (i.e., Avoid grammatical redundancy in the items.
satisfice) because of item Space related items apart.
characteristics or item
context
42 Podsakoff et al.
values their time, (e) personalizing the appeal by adding a signature from the researcher, and,
where possible, ( f ) promising feedback to participants so that they can gain self-awareness and
self-understanding (MacKenzie & Podsakoff 2012).
Of course, in addition to a lack of understanding of the importance of the information being
sought, respondents may lack the motivation to expend energy completing the survey either
because they are suspicious about the measurement context or because of the characteristics of
the items on the questionnaire or the context in which they are solicited. To lessen concerns
about the measurement context, researchers can explain (a) why the data are being gathered,
(b) that the data will be used only for research purposes, (c) that only aggregated information will
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
be provided to the organization, and (d) that individual responses will not be given to anyone
in the organization. Moreover, if identifying the respondent is not critical for other reasons
(e.g., matching responses from the survey with other information), the responses can be made
anonymous. Finally, to reduce the potential negative effects of item characteristics or item context
on respondents, researchers should, where possible, minimize the repetitiveness of the items as
well as the length of the questionnaire, if doing so does not compromise the construct validity of
the measures (MacKenzie & Podsakoff 2012).
Remedies directed at task difficulty. Another condition that has been found to be related to
CMB is that the task (survey) is too difficult, which may result from the use of ambiguous, dif-
ficult, or complex language. To address these conditions, researchers should (a) use clear and
concise language, (b) simplify complex or abstract concepts and questions, (c) clarify vague con-
cepts with examples, (d) avoid double-barreled questions, and (e) simplify questions and response
options.
Remedies directed at item characteristics and item contexts. Finally, given that item char-
acteristics and the context in which they are presented can affect the opportunity to satisfice,
researchers interested in reducing the effects of CMB can, where possible, (a) vary the types of
scales, scale points, and anchor points; (b) space related items apart; and (c) avoid grammatical
redundancy across the measures of the predictor and criterion variables (Cortina et al. 2020).
However, when modifying the properties of any scale, it is important to assess the construct va-
lidity of the adapted measures (for the types of evidence that can be used to support the validity
of adapted scales, see Heggestad et al. 2019).
HSF
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
M
UMLV
Figure 7
Illustrations of statistical remedies for controlling common method bias. (a) Basic CFA model. (b) CFA HSF model. (c) CFA-based
UMLV technique. (d) CFA-based DMLV approach. Abbreviations: CFA, confirmatory factor analysis; CMV, common method variance;
DMLV, directly measured latent variable; HSF, Harman’s single-factor technique; M, the specific method factor used in the analysis
(which could include a marker variable, a directly measured latent variable, or measured response styles); UMLV, unmeasured latent
variable technique.
Harman’s single-factor technique. The assumption of HSF is that if CMV is present and has a
confounding effect, it will manifest itself through the presence of a single dominant factor. Histor-
ically, the absence of a single dominant factor or the presence of additional factors in a model was
used to support the conclusion that CMV was not a problem. In terms of implementation, all the
items of the substantive variables included in a study are loaded into an exploratory factor analysis
(EFA). If the first unrotated factor does not account for the majority (>50%) of variance in the
items, it is inferred that CMB is not a problem. This EFA approach has recently been replaced or
supplemented with the CFA HSF model shown in Figure 7b, which excludes substantive latent
factors and replaces them with a single method factor that loads on all the substantive indica-
tors. Then the fit of the HSF model is examined using one or more indices for model evaluation,
and if it is determined to be worse than the originally proposed substantive model, CMV is not
considered to be of concern.
The basic advantage of both EFA and CFA HSF techniques is their ease of implementation.
Since they do not require any preplanning at the study design stage, nor do they require re-
searchers to include additional, nonfocal measures in their questionnaires, they are often used
in a post hoc fashion as a response to journal editor/reviewer concerns about CMB. However,
this convenience is substantially offset by the limitations of HSF. First, as noted by several re-
searchers (Aguirre-Urreta & Hu 2019, Baumgartner et al. 2021, Hulland et al. 2018, Jakobsen
& Jensen 2015), the single common factor extracted from EFA is likely to confound substantive
variance and CMV, leading to false positives when correlations between substantive constructs are
high and false negatives when correlations are medium to low. Relatedly, HSF is subject to low
44 Podsakoff et al.
statistical power because it can detect CMV only when the majority of the variance in the ratings of
the substantive constructs can be explained by one factor, making the technique very conservative
(Aguirre-Urreta & Hu 2019, Baumgartner & Weijters 2021, Baumgartner et al. 2021, Hulland
et al. 2018, Jakobsen & Jensen 2015, Schwarz et al. 2017, Steenkamp & Maydeu-Olivares 2021).
Third, HSF does not identify the specific sources of method variance affecting the focal relation-
ships, and it assumes a single source of method variance for all the items—which is doubtful, given
that more than one source of method variance is likely to be present in all survey studies. Finally,
the criterion used to determine how much variance the first factor should explain (e.g., >50%) is
arbitrary (Podsakoff et al. 2003, Steenkamp & Maydeu-Olivares 2021).
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Although it appears to be more rigorous, the CFA HSF approach suffers because if the hy-
pothesized measurement model fits the data better than a single-factor solution, this procedure
provides evidence of the fit of the measurement model and not evidence for the absence of CMB.
However, perhaps the biggest limitation of both EFA and CFA HSF techniques is that they do not
control for CMB—they simply try to determine whether this form of bias is present in the data by
using a very unsophisticated tool. This approach is problematic because even if the first unrotated
factor does not account for 50% of the variance in the items included in EFA, the presence of
shared method variance across the focal variables will produce biased estimates of the observed
relationships. Similarly, the rejection of the CFA-based HSF model does not mean that there is no
CMV or bias present that may be compromising researchers’ conclusions. Thus, like several other
researchers before us (Aguirre-Urreta & Hu 2019; Baumgartner & Weijters 2021; Baumgartner
et al. 2021; Hulland et al. 2018; Jakobsen & Jensen 2015; Podsakoff et al. 2003, 2012; Schwarz
et al. 2017; Steenkamp & Maydeu-Olivares 2021), we strongly recommend against using either
EFA or the CFA HSF technique.
Unmeasured latent variable technique. The CFA-based UMLV technique can be thought of
as taking the HSF and adding it to the original correlated substantive factor model (rather than
using it to replace the substantive latent variables). Figure 7c illustrates the resulting model. To
test for the presence of CMB, researchers compare (a) the fit of the model that includes only the
substantive factors and their loadings with (b) the fit of the model that includes both the substantive
and method variable factors and loadings. If the fit of the two models is not significantly different
(e.g., the chi-square difference is less than that of the critical values for its degrees of freedom),
then the researcher can conclude that CMV is not present (and infer that CMB is not of concern).
However, if the model that includes the unmeasured method factor significantly improves the
fit of the model, then it is important to control for these method effects to determine whether
the observed relationships between the substantive variables are still significant. If the empirical
relationships between the predictor and criterion variables remain significant after the inclusion
of the unmeasured latent method factor, researchers can conclude that their relationships hold
even after statistically controlling for sources represented by the unmeasured method factors.
The UMLV technique has several potential advantages. First, it is relatively easy to implement
after the primary data for the study have been acquired, and it does not require additional measures.
Second, the UMLV technique does not require researchers to identify or measure the specific
variable(s) responsible for the method effect(s). Third, this technique models the effect of the
method factor at the item level, rather than at the construct level (Podsakoff et al. 2003, 2012;
Williams et al. 1996). Finally, the UMLV technique does not require the effects of the method
factor on each indicator to be equal (Podsakoff et al. 2012).
However, the technique also has several limitations. First, as noted by several researchers
(Bagozzi 2011, Podsakoff et al. 2003, Williams & McGonagle 2016), one cannot be sure what
specific sources of CMV are being captured by the latent method factor, and the method factor
may reflect not only different types of method variance but also variance due to relationships
it “is neither able to detect, nor control for, common method bias. Method estimates using this
approach resulted in negligible estimates, regardless of whether there were some, large, or no
method bias introduced in the simulated data.” In addition, Chin et al. noted that the PLS-based
UMLV technique is problematic because, among other things, every indicator in the PLS model
includes not only variance attributable to the trait of interest but also method variance, and any
of the structural paths among the constructs will still be biased by method effects.
The UMLV approach suffers from technical challenges as well. For example, Podsakoff et al.
(2012) have noted that adding the method factor to the model can cause identification problems
if the ratio of the number of indicators to the number of substantive constructs is low. Identifi-
cation problems also are likely if the method factor loadings have similar or equal values (e.g.,
Kenny & Kashy 1992). Furthermore, this procedure assumes that the substantive variable fac-
tors do not interact with the method factor, an assumption that has been questioned by several
researchers (Bagozzi & Yi 1990, Campbell & O’Connell 1967, Podsakoff et al. 2003, Wothke &
Browne 1990). However, in our view the biggest problem with the UMLV approach is its use of
a single latent variable to represent what are very likely to be multidimensional sources of CMV.
This problem was illustrated in a recent study reported by Spector et al. (2019). After apply-
ing the UMLV technique to simulated data, these authors concluded that the “results indicate
that simplistically modeling one method factor in circumstances where method variance actually
stems from multiple method factors with (a) distinct patterns of relationships with one another and
(b) across substantive items can, in some cases, produce more erroneous results than does modeling
no method factors at all” (Spector et al. 2019, p. 866).
Marker variable technique. The next set of approaches used for detecting and controlling
CMV differ from HSF and the UMLV technique in that they incorporate measures of variables
presumed to reflect specific forms of method variance in the analyses. Figure 7d shows a general
model for these approaches. Note that, in contrast to the UMLV, this model includes indicators
used to capture the specific types of variance that may be the source of CMV/CMB. The first
CFA technique we present uses MVs that are presumed to share measurement characteristics with
the substantive constructs of interest but are conceptually unrelated to these variables. Lindell
& Whitney (2001) introduced a partial correlation technique for use of MVs that assumes that
any empirical relationship between the MV and a substantive variable represents method-related
biases. As noted by Richardson et al. (2009), an ideal MV should meet three criteria: It should
be (a) chosen a priori and (b) theoretically unrelated to the substantive variables but (c) similar
to them in content and format. Although the MV technique was originally designed to control
for CMB by partialling out the smallest correlation between the MV and a substantive variable
included in the study to determine whether the relationships between the substantive variables
were still significant, more recently researchers have used a more sophisticated regression-based
analysis (Siemsen et al. 2010). In the regression-based technique, the MV is added to the
regression equation, along with the predictors of the focal criterion variables, to determine
whether the substantive variables remain significantly different from zero.
46 Podsakoff et al.
Williams et al. (2010) proposed a CFA-based MV technique to overcome the limitations of
the partial correlation and regression procedures, which can be understood via Figure 7d. In
this approach, a series of measurement models representing the relationships between the latent
MV and the substantive indicator variables are compared under different constraints to determine
(a) whether the latent MV has (un)equal effects on the substantive variables and (b) whether the
relationships between the substantive variables are biased due to the latent MV (for a detailed
description, see Williams et al. 2010).
The CFA-based MV technique is clearly more sophisticated than either the correlation-based
or regression-based technique because these latter techniques (a) account for method variance
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
only at the construct level, not at the level of the indicators; (b) do not control for measurement
error; and (c) assume that CMV has equal effects on all observed variables, which is not supported
by the available evidence (Baumgartner & Steenkamp 2001, Cote & Buckley 1987, Williams et al.
2010). In addition, the correlation-based MV technique assumes that CMV can only inflate (and
not deflate) correlations among the substantive variables and that the regression-based technique
applies only to single-equation models. Nevertheless, all three of the MV techniques have im-
portant limitations. First, it is unclear what specific sources of CMV the MV techniques are
controlling (Simmering et al. 2015). Second, it is doubtful that any given MV will control for
all potential sources of CMB (Williams & McGonagle 2016). More importantly, it is unlikely
that CMB that is attributable to relationship-specific sources of CMB (e.g., implicit theories and
consistency motifs) is controlled using this technique, because the MV is supposed to be unre-
lated to the substantive variables included in a study. This limitation is important because implicit
theories have been shown to influence a variety of substantive relationships in the field (Eden &
Leviatan 1975, Lord et al. 1978, Smither et al. 1989), and there is considerable evidence that people
try to maintain consistency in their responses to events/survey items. Finally, several researchers
(Baumgartner & Weijters 2021, Steenkamp & Maydeu-Olivares 2021) have noted that the practice
of selecting MVs often does not satisfy the criteria identified by Richardson et al. (2009) and does
not include the emphasis on linking to measurement theory advocated by Williams et al. (2010).
Directly measured latent variable technique. The DMLV technique improves on the CFA
marker technique by using measures that are presumed to directly represent processes that gen-
erate CMV. Williams & Anderson (1994) and Williams et al. (1996) originally used the term
measured method effect variables in this context, and they delineated how the technique should
be implemented. More recently, Williams & McGonagle (2016) used the term measured cause
to refer to DMLVs. To apply the DMLV technique, researchers identify the specific source of
CMB that they believe might be present in a study, directly measure it, and control for it in the
analyses. As such, the form of the CFA model for the DMLV technique is the same as for the
CFA marker model shown in Figure 7d. The only difference is in the nature of the variables
used in the attempt to detect and control for CMV. The most common variables included by re-
searchers when using this technique include social desirability (Barrick & Mount 1996), positive
or negative affect (Schaubroeck et al. 1992, Williams & Anderson 1994, Williams et al. 1996), im-
pression management (Brady et al. 2017), and response styles (Baumgartner & Steenkamp 2001,
Weijters et al. 2008). Once the specific form of CMB that is expected to influence the substantive
relationships has been identified and measured, (a) the method effect indicators included in the
study are allowed to load on their theoretical constructs and on a latent method factor that has
its own measurement component, and (b) the significance of the factor correlations or structural
parameter estimates is examined both with and without the latent method factor in the model. If
the parameter estimates remain significant with the method factor included, the researcher can
conclude that the structural relationships are supported.
techniques discussed above, it is unlikely that this technique can be used to control for some
of the most potent causes of CMB (e.g., implicit theories and consistency motifs); however, for
exceptions, see the attempts by Phillips & Lord (1986) and Rush et al. (1977) to measure and
control for implicit leadership theories. With that said, we believe that the DMLV technique is
valuable when the potential sources of CMB can be identified and validly measured.
48 Podsakoff et al.
Table 5 Summary of future research directions
Future research direction Example research questions Example (or possible) studies
Experiments designed to How does item proximity affect CMV/CMB? Wilson et al. (2021) used an experimental
examine the effects of What are the influences of item characteristics, design to compare five item-ordering
method factors item contextual factors, rater characteristics, approaches: (a) individually randomized
and measurement context on CMV/CMB? items (randomized), (b) static items
What is the impact of transient mood on grouped by construct (grouped),
measures of work-related affect? (c) static intermixed items (intermixed),
(d) individually randomized grouped-
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Abbreviations: CMB, common method bias; CMV, common method variance; ESM, experience sampling method.
50 Podsakoff et al.
CMB inflates the relationships between CSE (a higher-order construct) and job satisfaction using
both statistical and procedural remedies. In terms of statistical remedies, these authors concluded
that controlling for a measured (e.g., social desirability) or an unmeasured latent method factor
appeared to be more effective at reducing CMB than controlling for an MV. In terms of pro-
cedural remedies, they found that temporal separation, particularly temporal separation of each
of the lower-level dimensions of the higher-order CSE construct, was more effective than us-
ing a combination of different response formats and filler scales between the subdimensions of
the higher-order CSE construct. However, it is unclear how many of these findings were due to
sampling differences, or whether they would generalize to other higher-order constructs. There-
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
fore, we encourage more research focusing on the effects of CMB on other multidimensional
constructs. Although several such constructs come to mind (overall job satisfaction, the Big Five
personality traits, challenge and hindrance stressors), given the important role that overall job per-
formance plays in the fields of management and applied psychology, and the fact that it is often
conceptualized as consisting of both positive (task performance and OCB) and negative (CWB)
subdimensions (Rotundo & Sackett 2002) that may cause several different types of CMB (e.g.,
trait and item social desirability/undesirability, implicit theories, leniency biases), it may prove to
be a particularly interesting construct to explore in future research.
Cross-Cultural Studies
Given the growing number of cross-cultural comparison studies that have been conducted during
the past 15 years (Broesch et al. 2020), and the fact that most of the studies reported in manage-
ment and applied psychology adapt questionnaires developed from English-speaking countries
(Harzing 2006, Harzing et al. 2012), perhaps it is not surprising that there is an increased interest
in the effects of cultural differences on CMB. However, it is surprising that most of the research in
this area has focused on how response styles differ across cultures (e.g., Baumgartner & Steenkamp
2001, Benitez et al. 2016, Harzing et al. 2012). Although we believe that such research is impor-
tant, we also believe that it is too restrictive in its scope and that there is a need for additional
research on the effects that cultural differences have on CMB. For example, the fact that cultural
differences have been found in socially desirable responding (Lalwani et al. 2009), implicit theories
(Church et al. 2012), and positive and negative affect (Bagozzi et al. 1999) suggests that additional
research should focus on the effects that culture has on CMB associated with rating sources.
CONCLUSION
It has been 30 years since Schmitt (1994) noted that in order to develop a better understanding
of the effects of CMB, researchers in the field of applied psychology need to make a stronger
commitment to (a) identify the sources of CMB, (b) clarify how these sources affect the nature of
substantively interesting relationships, and (c) explain how CMB could be measured and/or con-
trolled. We believe that the field has made considerable strides in addressing these issues (see the
Summary Points). For example, during the past few decades researchers have devoted significant
effort to identifying and categorizing the various sources of CMB (e.g., Baumgartner et al. 2021;
MacKenzie & Podsakoff 2012; Podsakoff et al. 2003, 2012; Weijters et al. 2009, 2010). Relatedly,
researchers have made a concerted attempt to uncover the mechanisms through which CMB has
its effects (MacKenzie & Podsakoff 2012, Podsakoff et al. 2012, Yao & Xu 2021) and the stages
in which CMB is likely to enter the survey response process (Podsakoff et al. 2003). We also have
a much better understanding of the effects that CMB can have on measures of construct validity
and reliability (Baumgartner & Steenkamp 2001, Cote & Buckley 1987, Williams et al. 2010) and
on the nature of the relationships between constructs (Baumgartner et al. 2021; Podsakoff et al.
2003, 2012; Williams et al. 2010). Moreover, various procedural and statistical remedies to control
CMB ( Jakobsen & Jensen 2015, MacKenzie & Podsakoff 2012) have been identified and imple-
mented, and there is growing evidence about which of these techniques work and which ones do
not (Baumgartner & Weijters 2021, Baumgartner et al. 2021, Hulland et al. 2018). This knowl-
edge should be gratifying to those researchers interested in understanding and trying to control
biases in their research. Thus, even though we have not resolved all the issues related to this
topic, we hope that we have provided a worthwhile summary of what we know about the sources
of CMB, its effects, the remedies that can be used to address it, and some directions for future
research.
52 Podsakoff et al.
SUMMARY POINTS
1. Despite recognition that method biases can have several harmful effects, the causes and
consequences of common method bias (CMB), and remedies for dealing with it, are still
not well understood.
2. Method factors are bad in that they can have harmful effects on estimates of construct
reliability and validity and on the empirical relationships between measures of different
constructs. CMB is more likely to be problematic when the rater lacks the ability or the
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
motivation to provide accurate ratings, the task (survey questionnaire) is difficult, and/or
raters are given the opportunity to exert low effort (i.e., satisfice).
3. CMB is complex for two reasons: It can result from several different sources, including
rater characteristics, item characteristics, item context effects, and the measurement con-
text, and it can inflate, deflate, or have no effect on the observed relationships between
focal constructs.
4. CMB appears to be widespread, as several disciplines, including applied psychology,
organizational behavior, marketing, operations management, and management informa-
tion systems, have reported that between 31% and 98% of published studies use designs
that are susceptible to it.
5. CMB is not easy to fix because many of the most commonly used statistical techniques
have substantial limitations, some remedies address one or only a few sources of potential
CMB but do not address other sources of CMB, and some techniques may mitigate one
or more sources of CMB while simultaneously magnifying the effects of other sources
of CMB.
6. The most widely used statistical remedy (i.e., Harman’s single-factor technique) is
limited and does not control for CMB, and we recommend strongly that its use be
discontinued.
7. When possible, researchers should use procedural remedies to control CMB. By obtain-
ing measures of the predictor and criterion variables from different sources and including
a temporal separation between them, researchers can effectively negate several of the
major causes of CMB.
8. Avenues for future research that should prove beneficial include experiments designed to
test the effects of different sources and remedies of CMB, remedies for CMB in measures
of multidimensional constructs, remedies for CMB in multilevel models and designs
using experience sampling methods, and comparisons of CMB effects across cultures.
DISCLOSURE STATEMENT
The authors are not aware of any affiliations, memberships, funding, or financial holdings that
might be perceived as affecting the objectivity of this review.
ACKNOWLEDGMENTS
P.M.P. and N.P.P. acknowledge the influence of Scott MacKenzie on this article. Although Scott
did not participate in its writing, his intellectual contributions will be evident to those readers who
are familiar with his previous collaborations with the first two authors on the topic of common
method bias.
Astra RL, Singg S. 2000. The role of self-esteem in affiliation. J. Psychol. 134:15–22
Bachrach DG, Bendoly E, Podsakoff PM. 2001. Attributions of the “causes” of group performance as an al-
ternative explanation of the relationship between organizational citizenship behavior and organizational
performance. J. Appl. Psychol. 86:1285–93
Bagozzi RP. 1984. A prospectus for theory construction in marketing. J. Mark. 48:11–29
Bagozzi RP. 2011. Measurement and meaning in information systems and organizational research:
methodological and philosophical foundations. MIS Q. 35:261–91
Bagozzi RP, Li YJ, Phillips LW. 1991. Assessing construct-validity in organizational research. Adm. Sci. Q.
36:421–58
Bagozzi RP, Wong N, Li YJ. 1999. The role of culture and gender in the relationship between positive and
negative affect. Cogn. Emot. 13:641–72
Bagozzi RP, Yi YJ. 1990. Assessing method variance in multitrait multimethod matrices: the case of
self-reported affect and perceptions at work. J. Appl. Psychol. 75:547–60
Barrick MR, Mount MK. 1996. Effects of impression management and self-deception on the predictive validity
of personality constructs. J. Appl. Psychol. 81:261–72
Baumgartner H, Steenkamp J. 2001. Response styles in marketing research: a cross-national investigation.
J. Mark. Res. 38:143–56
Baumgartner H, Weijters B. 2021. Dealing with common method variance in international marketing research.
J. Int. Mark. 29:7–22
Baumgartner H, Weijters B, Pieters R. 2021. The biasing effect of common method variance: some
clarifications. J. Acad. Mark. Sci. 49:221–35
Beal DJ, Weiss HM, Barros E, MacDermid SM. 2005. An episodic process model of affective influences on
performance. J. Appl. Psychol. 90:1054–68
Benitez I, He J, Van de Vijver FJR, Padilla J-L. 2016. Linking extreme response style to response processes: a
cross-cultural mixed methods approach. Int. J. Psychol. 51:464–73
Bliese PD. 2000. Within-group agreement, non-independence, and reliability: implications for data aggrega-
tion and analysis. In Multilevel Theory, Research, and Methods Inorganizations: Foundations, Extensions and
New Directions, ed. KJ Klein, SW Kozlowski, pp. 349–81. San Francisco, CA: Jossey-Bass
Bodner TE. 2006. Designs, participants, and measurement methods in psychological research. Can. Psychol.
47:263–72
Boehm SA, Kunze F, Bruch H. 2014. Spotlight on age-diversity climate: the impact of age-inclusive HR
practices on firm-level outcomes. Pers. Psychol. 67:667–704
Bollen KA. 1989. Structural Equations with Latent Variables. New York: Wiley
Bozionelos N, Simmering MJ. 2022. Methodological threat or myth? Evaluating the current state of evidence
on common method variance in human resource management research. Hum. Resour. Manag. J. 32:194–
215
Brady DL, Brown DJ, Liang LH. 2017. Moving beyond assumptions of deviance: the reconceptualization and
measurement of workplace gossip. J. Appl. Psychol. 102:1–25
Brannick MT, Chan D, Conway JM, Lance CE, Spector PE. 2010. What is method variance and how can we
cope with it? A panel discussion. Organ. Res. Methods 13:407–20
Broesch T, Crittenden AN, Beheim BA, Blackwell AD, Bunce JA, et al. 2020. Navigating cross-cultural
research: methodological and ethical considerations. Proc. R. Soc. B 287:20201245
Brown TA. 2015. Confirmatory Factor Analysis for Applied Research. New York: Guilford. 2nd ed.
54 Podsakoff et al.
Buckley MR, Cote JA, Comstock SM. 1990. Measurement errors in behavioral sciences: the case of
personality/attitude research. Educ. Psychol. Meas. 50:447–74
Burton-Jones A. 2009. Minimizing method bias through programmatic research. MIS Q. 33:445–71
Campbell DT, Fiske DW. 1959. Convergent and discriminant validation by the multitrait-multimethod
matrix. Psychol. Bull. 56:81–105
Campbell DT, O’Connell EJ. 1967. Methods factors in multitrait-multimethod matrices: multiplicative rather
than additive. Multivar. Behav. Res. 2:409–26
Chaffin D, Heidl R, Hollenbeck JR, Howe M, Yu A, et al. 2017. The promise and perils of wearable sensors
in organizational research. Organ. Res. Methods 20:3–31
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Chan D. 1998. Functional relations among constructs in the same content domain at different levels of analysis:
a typology of composition models. J. Appl. Psychol. 83:234–46
Chang SJ, van Witteloostuijn A, Eden L. 2010. From the editors: common method variance in international
business research. J. Int. Bus. Stud. 41:178–84
Chen PY, Dai T, Spector PE, Jex SM. 1997. Relation between negative affectivity and positive affectivity:
effects of judged desirability of scale items and respondents’ social desirability. J. Personal. Assess. 69:183–
98
Cheng KHC, Hui CH, Cascio WF. 2017. Leniency bias in performance ratings: the Big-Five correlates. Front.
Psychol. 8:521
Chin WW, Thatcher JB, Wright R. 2012. Assessing common method bias: problems with the ULMC
technique. MIS Q. 36:1003–19
Church AT, Willmore SL, Anderson AT, Ochiai M, Porter N, et al. 2012. Cultural differences in implicit the-
ories and self-perceptions of traitedness: replication and extension with alternative measurement formats
and cultural dimensions. J. Cross-Cult. Psychol. 43:1268–96
Chyung AY, Barkin JR, Shamsy JA. 2018. Evidence-based survey design: the use of negatively worded items
in surveys. Perform. Improv. 57:16–25
Colquitt JA, LePine JA, Zapata CP, Wild RE. 2011. Trust in typical and high-reliability contexts: building and
reacting to trust among firefighters. Acad. Manag. J. 54:999–1015
Connolly JJ, Viswesvaran C. 2000. The role of affectivity in job satisfaction: a meta-analysis. Personal. Individ.
Differ. 29:265–81
Cooper B, Eva N, Fazlelahi FZ, Newman A, Lee A, Obschonka M. 2020. Addressing common method variance
and endogeneity in vocational behavior research: a review of the literature and suggestions for future
research. J. Vocat. Behav. 121:103472
Cortina JM, Sheng Z, Keener KK, Keeler KR, Grubb LH, et al. 2020. From alpha to omega and beyond!
A look at the past, present, and (possible) future of psychometric soundness in the Journal of Applied
Psychology. J. Appl. Psychol. 105:1351–81
Cote JA, Buckley MR. 1987. Estimating trait, method, and error variance: generalizing across 70 construct-
validation studies. J. Mark. Res. 24:315–18
Cote JA, Buckley MR. 1988. Measurement error and theory testing in consumer research: an illustration of
the importance of construct-validation. J. Consum. Res. 14:579–82
Craighead CW, Ketchen DJ, Dunn KS, Hult GTM. 2011. Addressing common method variance: guidelines
for survey research on information technology, operations, and supply chain management. IEEE Trans.
Eng. Manag. 58:578–88
Cram WA, D’Arcy J, Proudfoot JG. 2019. Seeing the forest and the trees: a meta-analysis of the antecedents
to information security policy compliance. MIS Q. 43:525–54
Crampton SM, Wagner JA. 1994. Percept-percept inflation in microorganizational research: an investigation
of prevalence and effect. J. Appl. Psychol. 79:67–76
Cruz KS. 2022. Are you asking the correct person (hint: Oftentimes you are not!)? Stop worrying about un-
founded common method bias arguments and start using my guide to make better decisions of when to
use self- and other-reports. Group Organ. Manag. 47:920–27
Cui TX, Kam CCS, Cheng EH, Ho MY. 2022. Distinguishing between trait desirability and item desirability
in predicting item scores: Is informant evaluation of personality free from social desirability? Personal.
Individ. Differ. 196:111708
56 Podsakoff et al.
Huang C. 2013. Relation between self-esteem and socially desirable responding and the role of socially
desirable responding in the relation between self-esteem and performance. Eur. J. Psychol. Educ. 28:663–83
Hulland J, Baumgartner H, Smith KM. 2018. Marketing survey research best practices: evidence and
recommendations from a review of JAMS articles. J. Acad. Mark. Sci. 46:92–108
Ibrahim AM. 2001. Differential responding to positive and negative items: the case of a negative item in a
questionnaire for course and faculty evaluation. Psychol. Rep. 88:497–500
Jakobsen M, Jensen R. 2015. Common method bias in public management studies. Int. Public Manag. J. 18:3–30
Janiszewski C, Wyer RS Jr. 2014. Content and process priming: a review. J. Consum. Psychol. 24:96–118
Johnson RE, Rosen CC, Djurdjevic E. 2011. Assessing the impact of common method variance on higher
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
58 Podsakoff et al.
Podsakoff NP, Spoelma TM, Chawla N, Gabriel AS. 2019. What predicts within-person variance in applied
psychology constructs? An empirical examination. J. Appl. Psychol. 101:727–54
Podsakoff NP, Whiting SW, Welsh DT, Mai KM. 2013. Surveying for “artifacts”: the susceptibility of the
OCB–performance evaluation relationship to common rater, item, and measurement context effects.
J. Appl. Psychol. 98:863–74
Podsakoff PM, MacKenzie SB, Lee JY, Podsakoff NP. 2003. Common method biases in behavioral research:
a critical review of the literature and recommended remedies. J. Appl. Psychol. 88:879–903
Podsakoff PM, MacKenzie SB, Podsakoff NP. 2012. Sources of method bias in social science research and
recommendations on how to control it. Annu. Rev. Psychol. 63:539–69
Podsakoff PM, Organ DW. 1986. Self-reports in organizational research: problems and prospects. J. Manag.
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
12:531–44
Pujol-Cols L, Lazzaro-Salazar M. 2021. Ten years of research on psychosocial risks, health, and performance in
Latin America: a comprehensive systematic review and research agenda. J. Work Organ. Psychol. 37:187–
202
Rafferty AE, Griffin MA. 2004. Dimensions of transformational leadership: conceptual and empirical
extensions. Leadersh. Q. 15:329–54
Rafferty AE, Griffin MA. 2006. Refining individualized consideration: distinguishing developmental
leadership and supportive leadership. J. Occup. Organ. Psychol. 79:37–61
Richardson HA, Simmering MJ, Sturman MC. 2009. A tale of three perspectives: examining post hoc statistical
techniques for detection and correction of common method variance. Organ. Res. Methods 12:762–800
Roth PL, BeVier CA. 1998. Response rates in HRM/OB survey research: norms and correlates, 1990–1994.
J. Manag. 24:97–117
Rotundo M, Sackett PR. 2002. The relative importance of task, citizenship, and counterproductive
performance to global ratings of job performance: a policy-capturing approach. J. Appl. Psychol. 87:66–80
Rush MC, Thomas JC, Lord RG. 1977. Implicit leadership theory: a potential threat to the internal validity
of leader behavior questionnaires. Organ. Behav. Hum. Perform. 20:93–110
Sackett PR, Larson JR Jr. 1990. Research strategies and tactics in industrial and organizational psychol-
ogy. In Handbook of Industrial and Organizational Psychology, ed. MD Dunnette, LM Hough, pp. 419–89.
Palo Alto, CA: Consult. Psychol.
Schaubroeck J, Ganster DC, Fox ML. 1992. Dispositional affect and work-related stress. J. Appl. Psychol.
77:322–35
Schmitt N. 1994. Method bias: the importance of theory and measurement. J. Organ. Behav. 15:393–98
Schmitt N, Stults DM. 1986. Methodology review: analysis of multitrait-multimethod matrices. Appl. Psychol.
Meas. 10:1–22
Schriesheim CA. 1981a. The effect of grouping and randomizing items on leniency response bias. Educ. Psychol.
Meas. 41:401–11
Schriesheim CA. 1981b. Leniency effects on convergent and discriminant validity for grouped questionnaire
items: a further investigation. Educ. Psychol. Meas. 41:1093–99
Schulte M, Ostroff C, Shmulyian S, Kinicki A. 2009. Organizational climate configurations: relationships to
collective attitudes, customer satisfaction, and financial performance. J. Appl. Psychol. 94:618–34
Schwarz A, Rizzuto T, Carraher-Wolverton C, Roldan JL, Barrera-Barrera R. 2017. Examining the impact
and detection of the “Urban Legend” of common method bias. Database Adv. Inf. Syst. 48:93–119
Scott BA, Lennard AC, Mitchell RL, Johnson RE. 2020. Emotions naturally and laboriously expressed:
antecedents, consequences, and the role of valence. Pers. Psychol. 763:587–613
Sessions H, Nahrgang JD, Newton DW, Chamberlin M. 2020. I’m tired of listening: the effects of supervisor
appraisals of group voice on supervisor emotional exhaustion and performance. J. Appl. Psychol. 105:619–
36
Shipp AJ, Cole MS. 2015. Time in individual-level organizational studies: What is it, how is it used, and why
isn’t it exploited more often? Annu. Rev. Organ. Psychol. Organ. Behav. 2:237–60
Shipp AJ, Jansen KJ. 2021. The “other” time: a review of the subjective experience of time in organizations.
Acad. Manag. Annu. 15:299–334
Siemsen E, Roth A, Oliveira P. 2010. Common method bias in regression models with linear, quadratic, and
interaction effects. Organ. Res. Methods 13:456–76
9:221–32
Spector PE. 2019. Do not cross me: optimizing the use of cross-sectional designs. J. Bus. Psychol. 34:125–37
Spector PE, Bauer JA, Fox S. 2010. Measurement artifacts in the assessment of counterproductive work be-
havior and organizational citizenship behavior: Do we know what we think we know? J. Appl. Psychol.
95:781–90
Spector PE, Brannick MT. 2010. Common method issues: an introduction to the feature topic in Organizational
Research Methods. Organ. Res. Methods 13:403–6
Spector PE, Gray CE, Rosen CC. 2022. Are biasing factors idiosyncratic to measures? A comparison of
interpersonal conflict, organizational constraints, and workload. J. Bus. Psychol. 38:983–1002
Spector PE, Nixon AE. 2019. How often do I agree: an experimental test of item format method variance in
stress measures. Occup. Health Sci. 3:125–43
Spector PE, Pindek S. 2016. The future of research methods in work and occupational health psychology.
Appl. Psychol. 65:412–31
Spector PE, Rosen CC, Richardson HA, Williams LJ, Johnson RE. 2019. A new perspective on method
variance: a measure-centric approach. J. Manag. 45:855–80
Staw BM. 1975. Attribution of the “causes” of performance: a general alternative interpretation of cross-
sectional research on organizations. Organ. Behav. Hum. Perform. 13:414–32
Steenkamp JBEM, Maydeu-Olivares A. 2021. An updated paradigm for evaluating measurement invariance
incorporating common method variance and its assessment. J. Acad. Mark. Sci. 49:5–29
Thomas KW, Kilmann RH. 1975. The social desirability variable in organizational research: alternative
explanation for reported findings. Acad. Manag. J. 18:741–52
Thoresen CJ, Kaplan SA. Barsky AP, Warren CR, de Chermont K. 2003. The affective underpinnings of job
perceptions and attitudes: a meta-analytic review and integration. Psychol. Bull. 129:914–45
Tourangeau R, Rasinski K, D’Andrade R. 1991. Attitude structure and belief accessibility. J. Exp. Soc. Psychol.
27:48–75
Wang TY, Bansal P. 2012. Social responsibility in new ventures: profiting from a long-term orientation. Strateg.
Manag. J. 33:1135–53
Weijters B, Baumgartner H. 2012. Misresponse to reversed and negated items in surveys: a review. J. Mark.
Res. 49:737–47
Weijters B, De Beuckler A, Baumgartner H. 2014. Discriminant validity where there should be none:
positioning same-scale items in separated blocks of a questionnaire. Appl. Psychol. Meas. 38:450–63
Weijters B, Geuens M, Schillewaert N. 2009. The proximity effect: the role of inter-item distance on reverse-
item bias. Int. J. Res. Mark. 26:2–12
Weijters B, Geuens M, Schillewaert N. 2010. The individual consistency of acquiescence and extreme response
style in self-report questionnaires. Appl. Psychol. Meas. 34:105–21
Weijters B, Schillewaert N, Geuens M. 2008. Assessing response styles across modes of data collection. J. Acad.
Mark. Sci. 36:409–22
Williams LJ, Anderson SE. 1994. An alternative approach to method effects by using latent-variable models:
applications in organizational-behavior research. J. Appl. Psychol. 79:323–31
Williams LJ, Cote JA, Buckley MR. 1989. Lack of method variance in self-reported affect and perceptions at
work: reality or artifact. J. Appl. Psychol. 74:462–68
Williams LJ, Gavin MB, Williams ML. 1996. Measurement and nonmeasurement processes with negative
affectivity and employee attitudes. J. Appl. Psychol. 81:88–101
60 Podsakoff et al.
Williams LJ, Hartman N, Cavazotte F. 2010. Method variance and marker variables: a review and
comprehensive CFA marker technique. Organ. Res. Methods 13:477–514
Williams LJ, McGonagle AK. 2016. Four research designs and a comprehensive analysis strategy for in-
vestigating common method variance with self-report measures using latent variables. J. Bus. Psychol.
31:339–59
Wilson V, Srite M, Loiacono E. 2021. The effects of item ordering on reproducibility in information systems
online survey research. Commun. Assoc. Inf. Syst. 49:41
Wothke W, Browne MW. 1990. The direct-product model for the MTMM matrix parameterized as a 2nd-
order factor-analysis model. Psychometrika 55:255–62
Downloaded from [Link]. Guest (guest) IP: [Link] On: Sun, 03 May 2026 17:38:56
Yao MH, Xu YJ. 2021. Method bias mechanisms and procedural remedies. Sociol. Methods Res. In press. https://
[Link]/10.1177/00491241211043141
Zhang W, Yuan G, Xue R, Han Y, Taylor JE. 2022. Mitigating common method bias in construction
engineering and management research. J. Constr. Eng. Manag. 148:0402208
Zeng B, Wen H, Zhang J. 2020. How does the valence of wording affect features of a scale? The method effects
in the undergraduate learning burnout scale. Front. Psychol. 11:585179
Zhu D, Doan T, Kanjanakan P, Kim PB. 2022. The impact of emotional intelligence on hospitality employees’
work outcomes: a systematic and meta-analytical review. J. Hosp. Mark. Manag. 31:326–47
Zohar D, Luria G. 2004. Climate as a social-cognitive construction of supervisor safety practices: scripts as
proxy of behavior patterns. J. Appl. Psychol. 89:322–33