Measuring Variables
Measuring Variables
Learning Objectives
Chapter Outline
a. Variable
* b. Measurement
c. Validity
d. Reliability
8. ______________ refers to the consistency of results and ____________ is the extent to which
you are measuring what you think you are measuring.
a. Reliability; periodicity
b. Validity; reliability
* c. Reliability; validity
d. Convergence; divergence
10. In order to establish the reliability of a measure of intelligence, Kevin administers two forms
of the test to a group of students. Which of the following reliability coefficient values would
indicate the most reliability for the test?
a. 0.35
* b. 0.85
c. -0.85
d. 2.20
11. Jenna would like to establish the reliability of a new measure of self-esteem but she doesn’t
have enough time to administer her test more than once. Which of the following methods of
establishing reliability would you suggest to Jenna?
a. Equivalent forms
* b. Internal consistency
c. Multidimensional
d. Concurrent
12. If we include items assessing memory, logic, and verbal comprehension on an intelligence
test – as opposed to food preferences or shoe size – then we have satisfied which of the following
types of validity?
a. Discriminant
b. Convergent
* c. Face
d. Internal
18. Suppose you have created a new method of diagnosing anxiety disorders. How could you
demonstrate that your method is construct valid?
a. teach several licensed clinicians to use and evaluate your method.
b. use your method to diagnose the same group of participants repeatedly over the course
of several years, and see if you consistently arrive at the same diagnosis for a given
individual
c. have several specialists read about your new technique and invite their opinions
* d. use your method to diagnose a group of participants, then see if your diagnoses match
with diagnoses taken from other, established methods
19. Tom wanted to assess the reliability of his measure of anxiety so he had a group of
introductory psychology students complete the measure of anxiety on March 3rd and again on
March 25th. He then compared the scores that the students made on the two testing occasions
using a statistical technique called correlation. He used this quantitative index as his measure of
reliability. Tom used what method to assess reliability?
* a. test-retest
b. equivalent forms
c. split-half
d. Cronbach’s alpha
20. Eduardo decided to assess the reliability of the carbohydrate craving inventory he
constructed. He had constructed two identical versions of the inventory and a group of 50 people
took both versions. Then Eduardo compared the responses of these 50 people on the two
versions of the craving inventory for his assessment of the reliability of the inventory. Eduardo
used what method to assess reliability?
a. test-retest
* b. equivalent forms
c. split-half
d. Cronbach’s alpha
21. Jacqueline wanted to assess the reliability of ratings made of children’s aggressive behavior
so she had two students rate the degree of aggression displayed by each of 50 children while
engaged in play. She then compared the ratings made by these two students and computed the
degree of agreement between them. Jacqueline used what method of assessing reliability?
a. split-half reliability
* b. interrater reliability
c. internal consistency reliability
d. test-retest reliability
22. Afiya wanted to assess the reliability of students’ observations of children’s aggressive
behavior so she had two students observe 100 behaviors displayed by each of 50 children while
engaged in play. After viewing each behavior the students recorded the behavior as being
aggressive or nonaggresive. Afiya then computed the percentage of times the two students
agreed on their assessment of each behavior. Afiya used what method of assessing reliability?
* a. interobserver agreement
b. Cronbach’s alpha
c. internal consistency reliability
d. test-retest reliability
23. Owen wanted to assess the internal consistency of his measure of anxiety so he measured
reliability estimates by comparing items within his test. He reported a reliability estimate of .80.
What measure of reliability is Owen using?
* a. Cronbach’s alpha
b. path analysis
c. factor analysis
d. Riesen’s reliability estimate
24. Dr. Smarsh creates a new intelligence test and wants to assess its reliablility. She administers
her test to the same sample of individuals on two separate occasions one week apart. What type
of reliability is Dr. Smarsh using?
25. The publisher of the SAT can assess this type of reliability when individuals take the SAT
more than one time with different versions of the SAT being given each time. Which type of
reliability can be assessed?
* a. percentage of agreement
b. correlation coefficient
c. t-test
d. coefficient alpha
27. Gerald is developing a measure of shyness and he determines that students scoring high on
the measure also score high for introversion on a well-known introversion-extraversion scale.
The outcome best illustrates
a. face validity.
* b. concurrent validity.
c. predictive validity.
d. discriminant validity.
29. Students sometimes complain that scores on the Graduate Record Exam (GRE) are not
related to how well students perform in graduate school. Essentially the students are saying
that the GRE does not have
a. reliability.
b. internal consistency.
* c. predictive validity.
d. discriminant validity.
32. One way to assess construct validity is to establish that scores on the test in question do NOT
correlate with established scales that are dissimilar or conceptually unrelated concepts. (e.g, a
scale to measure depression would likely not correlate with scales designed to measure
happiness). This type of validity is called
a. concurrent validity.
b. convergent validity.
* c. discriminant validity.
d. predictive validity.
33. A study examines scores on an employment test and job performance six months later. This
study is most likely attempting to establish
* a. criterion validity.
b. face validity.
c. reliability.
d. construct validity.
34. Construct validity is most related to what other concept discussed in research methods?
a. reliability
* b. operational definitions
c. falsifiability
d. measurement
a. Predictive
b. Discriminant
c. internal
* d. Content
36. The extent to which groups that are known to diverge from one another actually differ on the
construct being developed is known as
a. discriminant validity
b. predictive validity evidence
* c. known groups validity evidence
d. criterion validity
37. When evaluating reliability and validity information of a published measure, it is important
to note the __________ upon which the information was gathered.
a. predictive sample
* b. norming group
c. estimate group
d. peer group
38. Some psychological tests are designed to measure more than one construct, they are
multidimensional. __________ is a statistical technique that can be used to determine the number
of dimensions that a particular measure is testing.
a. Cronbach’s alpha
c. Path analysis
c. Operationalization
* d. Factor analysis
40. ___________ refers to any sampling method in which each individual has an equal chance of
being selected for the sample.
* a. Equal probability selection method (EPSEM)
b. Convenience sampling
c. Stratified sampling
d. Homogeneous sampling selection
42. A ____________ is the full set of all individuals of interest and is typically hard to assess
fully.
a. sample
* b. population
c. sampling error
d. statistic
43. Dr. Saucer is in the process of drawing individuals from a population to be included in a
sample. This process is called
a. sampling frame
b. parameter
* c. sampling
d. proximal similarity
44. Dr. Konrad contacts the registrar at his university to get a complete list of students enrolled.
He wants to draw a simple random sample from the entire population. He does this to make sure
he is creating a
a. framed sample
* b. representative sample
c. stratified sample
d. cluster sample
45. When using simple random sampling it is suggested that you _________ because it will lead
to a more representative sample.
a. sample with replacement
* b. sample without replacement
c. sample half with replacement and half without replacement
d. use a larger sample than if using a nonrandom method
48. Suppose you wish to test a representative sample of people in your psychology class on
attitudes toward on-line courses. There are 40 people in the class, 30 females and 10 males.
What would be the most efficient strategy to ensure that your sample reflects the distribution
of males and females in the classroom population?
a. simple random sample
b. cluster sample
* c. proportional stratified sample
d. convenience sample
51. Although statistics are calculated to represent parameters, we know that there can be a
difference between the calculated statistic and population parameters. This difference is called
* a. sampling error
b. sampling frame
c. sampling difference
d. parameter difference
52. Dr. Monroe is interested in surveying college students at her university who are first
generation college students. She contacts the registrar’s office and finds out that there are 3500
first generation college students total. She wants to include 350 of those students in her sample.
She then divided the population by the sample size and gets k = 10. She then picks a random
number between 1 and 10 which is 8. This determines that every 8th person in the population will
be included in the sample. What type of sampling method is this?
a. cluster sampling.
b. stratified sampling
c. random sampling
* d. systematic sampling
53. To study career aspirations among high school students in Alabama, a researcher randomly
selects 5% of the state’s school districts and gives all the students in each district a survey
designed to measure career goals. What sampling procedure is being used here?
a. quota
b. stratified
* c. cluster
d. EPSEM
a. random
b. representative
* c. biased
d. systematic
55. Jim is conducting a survey to learn about student attitudes toward abortion. He passes out his
survey to the first 100 students that enter the cafeteria. What sampling method is Jim using?
* a. convenience
b. cluster
c. stratified
d. EPSEM
57. Which of the following nonrandom sampling techniques would be most similar to stratified
random sampling?
a. convenience
* b. quota
c. cluster
d. purposive
58. If you are interested in participants in your study with very specific characteristics (e.g.
English as a second language) you may conduct ________ sampling and once you contact some
individuals who want to participate, you may then conduct ________ sampling in which
participants identify other potential participants.
a. snowball; purposive
b. matched; snowball
c. random; non-random
* d. purposive; snowball
61. Your textbook authors suggest that if your population has fewer than 100 people you should
a. use a stratified sampling method.
b. use simple random selection.
c. find a larger population
* d. test everyone in the population.
62. Which of the following situations would NOT necessitate a larger sample size?
a. if your population is heterogeneous
b. if you plan to use multiple categories
c. if you expect a weak effect
* d. if you use proportional stratified sampling
63. Which of the following was NOT offered as a type of sampling method used for qualitative
research?
* a. purposive sampling
b. extreme case sampling
c. homogeneous sample selection
d. maximum variation sampling
Vocabulary
Essay questions
1) Define variable and measurement. Describe the difference between them and give an
example for each.
2) List the four scales of measurement from least complex to most complex – provide an
example of each.
4) Using an example, explain the difference between reliability and validity. Explain
what it means to say that a reliable measure is not always valid, but a valid measure is
always reliable.
8) Discuss content validity and criterion-related validity (be sure to include a discussion
of predictive and concurrent validity).
9) What is factor analysis and how would the results be different if the assessment tool
was uni-dimentional compared to multi-dimensional?
10) Assume that you have created a new test to measure depression. Explain how you
could use converging evidence and discriminant evidence to establish the validity of your
instrument.
11) What is known groups validity evidence? Give an example to help explain your
answer.
12) Describe, in general terms, the difference between random sampling methods and
nonrandom sampling methods. Briefly describe the four random sampling techniques
presented in the text.
13) Describe the four nonrandom sampling techniques presented in your text. If, as
indicated in your text, these techniques are “weaker sampling methods” why would
researchers use them?
14) Discuss the factors that are important for researchers to consider when determining
sample size.
15) Briefly summarize some of the sampling methods used in qualitative research.
1) Your students may already be familiar with the four scales of measurement but a simple way
to personalize your lecture is to provide them with a brief (anonymous) survey to complete at the
beginning of class. A sample survey is included below that includes examples of all four
measurement scales. After students complete the survey ask them to identify the items
representing each scale of measurement. You might also consider collecting the surveys and
using the data later in the semester to illustrate some simple statistical techniques.
1 2 3 4 5 6 7
Very Neutral Very
Negative Positive
2) To help students understand the concepts of reliability and validity, you might consider the
exercise proposed by Miserandino (2006) who used an Internet-based test as a fun way to
reinforce these topics. This assignment has the added benefit of reminding students, once again,
of the importance of critically evaluating information that they encounter on the Internet.
Miserandino, M. (2006). I scream, you scream: Teaching validity and reliability via the
ice cream personality test. Teaching of Psychology, 33 (4), 265-268.
3) Some of your students may believe that psychological research isn’t valid unless research
participants are randomly selected. Students should be aware that because of large populations of
interest, it may not be feasible to use random selection. Furthermore, you should point out that
because much research in psychology is “basic” – attempting to establish general relationships
between variables – random selection is less important. On the other hand, random assignment is
crucial in helping to eliminate extraneous variables and allowing an unambiguous interpretation
of our results. A good discussion of this distinction can be found in the Stanovich (2009) book
referenced below. The Enders et al. (2006) article describes a very simple classroom
demonstration of random assignment using a deck of playing cards.
Enders, C., Laurenceau, J., & Stuetzle, R. (2006). Teaching random assignment: A
classroom demonstration using a deck of playing cards. Teaching of
Psychology, 33(4), 239-242.
Stanovich, K.E. (2009). “But it’s not like real life!” The artificiality criticism and
psychology. In How to Think Straight About Psychology. Boston, MA: Allyn &
Bacon.
4) The Web site Clips for Class ([Link]) has an extensive list of videos for use
in many different psychology classes. Under the Statistics tab there are several videos that could
help students in this chapter. They are listed below with descriptions found on the Web site and
suggested questions for students after watching each clip.
Nominal Variables Video
This video explains nominal variables. How do these differ from ordinal variables?
Sesame Street: Neil Patrick Harris Has Telly’s New Shoes
In this brief song, characters sing and dance about shoes. Is the list of shoes given a
nominal scale, ordinal scale, interval scale, or ratio scale? What kind of descriptions
of shoes could be given to change it to a different kind of scale?
5) The Web site Clips for Class ([Link]) has an extensive list of videos for use
in many different psychology classes. Under the Research tab there are several videos that could
help students in this chapter. They are listed below with descriptions found on the Web site and
suggested questions for students after watching each clip.
Reliability
The headless professor discusses the concept of reliability and some of the more
common types of reliability measures. He presents a 2 x 2 contingency table to
demonstrate reliability as having more agreements than disagreements. What do
measures of reliability determine? Why did the headless professor present the 2 x 2
table to explain reliability? How is reliability defined in that instance? Why is a type
of interrater reliability important for clinicians?
Validity
The headless professor defines validity and how it is established by way of a 2 x 2
contingency table. He also discusses different types of validity such that in clinical
psychology (diagnosis and prognosis) and industrial/organizational psychology
(performance). Define validity in your own words. Why is it important to establish
validity in clinical psychology and industrial/organizational psychology? Why did the
headless professor present the 2 x 2 table to explain validity?
Samples
The headless professor explains the difference between subjects (participants),
groups, samples, and populations. He offers a clinical example and a marketing one,
and provides a diagram showing the relationship between a population, group,
sample, and participants. What is the difference between subjects/participants,
groups, samples, and populations? Develop a research question. Are you able to
identify your participants, relevant groups, secure a sample, and estimate your
population of interest?
6) Raechel Soicher created several videos demonstrating different sampling techniques using
Jing. Here are the links to those videos.