Research Methods Notes 2
Research Methods Notes 2
The Independent Variable (IV) is the variable that is manipulated by the researcher to
check its effect on the dependent variable.
The Dependent Variable (DV) is the variable being measured by the researcher.
Research Methods
Experiments
Experiments look for a causal relationship in which an independent variable is
manipulated to cause a change in the dependent variable.
Experiment types:
Field experiment – the experiment takes place in natural settings, and the IV is
manipulated by the researcher.
Natural experiment – the experiment takes place in a natural setting and the IV
is not directly manipulated by the researcher. It happens naturally by chance.
Quasi-experiment – the researcher has lots of control over the procedure, but
not over the allocation of participants.
Weaknesses:
Since the experiment takes place in artificial settings, there are chances of
participants exhibiting demand characteristics and social desirability bias.
Participants could act on the demand characteristics as a response and alter their
behaviour which would lead to lower validity and ecological validity.
Furthermore, the mundane realism of any tasks involved in the lab exp. would be
low. Mundane realism is the extent to which any tasks involved in a study
represent real-world activities.
Participants are in natural settings showing their natural behaviour, hence they
are likely to show their true behaviour. This would make the results highly valid,
ecologically valid, and representative.
If participants are unaware that they are in a study, there is very less chance of
participants exhibiting demand characteristics, in comparison to lab experiments.
Weaknesses:
The researcher will be less sure that changes in the DV are caused by changes in
the IV, in comparison to lab experiments. This leads to lower validity.
The participants are unaware that they are in a study, and this raises the ethical
issues of privacy, confidentiality, and lack of consent.
When participants are in natural settings, their behaviour is likely to be true, and
representative. This means the ecological validity is high.
As participants are unaware of their participation in the research, there are very
less chances of demand characteristics affecting the experiment.
Natural experiments are only possible to conduct when changes in variables occur
naturally.
As the researcher isn’t manipulating the IV, they will be less sure of the cause of
changes in the DV, so a causal relationship cannot necessarily be established.
They are hard to replicate as controls and standardization are hard to implement,
so the reliability may be low.
Experimental Design: how participants are allocated to the conditions of the study.
Experimental Design types:
Independent measures design – different groups of participants are used for each
level of the IV.
Repeated measures design – each participant takes part in every condition of the
study.
Order effects occur when participants take part in more than one condition of
the study.
Use counterbalancing to overcome the order effects from impacting the study’s results,
when implementing the repeated measures design.
When implementing the repeated measures design in a study, there are chances of order
effects impacting the study’s results. Therefore, counterbalancing can be used to reduce
order effects.
Counterbalancing – When the order in which each group attends each level of the IV is
the opposite. A,B. B,A.
How to improve generalisability: Use randomization to select the sample for the
study.
Randomisation: when you choose your sample for the study randomly.
Self Reports
1. Questionnaires – research method using written questions.
Question types:
o Close-ended (pre-set answer choices, quantitative data)
Evaluation:
0. Easier to analyse than interviews as it isn’t affected by researcher bias when
calculating results.
b. Responses may be limited – low validity as they can’t explain the reason.
Evaluation:
a. Difficult to quantify.
b. There is a chance that researcher bias might affect the interpretation of results.
General Questionnaire Evaluations:
Easy for participants to ignore questions – low generalisability.
Case Studies are detailed investigations about a single person or a small group. The
maximum amount of qualitative and quantitative data is gathered.
Strengths:
The data collected is highly valid.
The researcher builds rapport with the subject, making it likely for them to open
up and provide true information.
The subject is less likely to show demand characteristics as case studies are
longitudinal studies.
Weaknesses:
The researcher’s findings may be biased due to the close relation with the
subject.
Observations
Naturalistic & controlled observations
An observation is a non-experimental method which
involves observing and recording behaviours in either naturalistic or
controlled settings
o the cult discuss the impending apocalypse because they have all been
brainwashed
All that can be recorded is the action/behaviour which is then linked to the topic of
the investigation with no assumption of cause-effect
Naturalistic observation
A naturalistic observation is one in which the researcher observes and records
behaviours in a natural setting, away from the lab, with no manipulation or a
complete absence of an independent variable (IV)) e.g.
If the naturalistic observation is covert, participants are unaware they are being
observed and are therefore unlikely to experience the Hawthorne effect
o This is the tendency to change behaviour simply because one knows they
are being studied
Limitations
Naturalistic observations cannot be easily replicated due to the uncontrolled
nature of the setting
o If the observation is overt, these ethical concerns are reduced but the risk
of demand characteristics increases
Controlled observation
A controlled observation is one in which the researcher implements a level of
control, implementing replicable procedures and (sometimes) an IV
separation anxiety
stranger anxiety
reunion behaviours
Participants know that they are taking part in a controlled observation as they
must be recruited for the study and then set a specific task which is likely to be
quite removed from their everyday activities/experience
o Thus this method has good reliability, particularly if more than one
observer is used throughout (known as inter-observer reliability)
o In Bandura's study only the children who had observed the aggressive
model performed imitative acts on the Bobo doll
o This supports the validity of the researcher's hypothesis
Limitations
The use of controlled conditions and artificial tasks means that controlled
observations are low in ecological validity
o This means that both mother and baby may have been responding in ways
which did not truly represent their attachment style
o The children in Bandura's study may simply have been aggressive because
they thought that this was expected of them (as they had seen an adult
behaving aggressively)
o This would lower the validity of the findings as it would not be a true
effect of the IV on the DV
o participants are not aware that they are being observed and will not have
been informed of this in advance
o shoppers in a mall
Limitations
There are ethical issues with covert observations
o In Piliavin's New York subway study, the passengers were deceived into
thinking that someone had collapsed in their carriage, which could have
caused them great distress
o Piliavin's study could not be replicated not only due to ethics but for
the very sound reason that anyone acting suspiciously on public
transport in the 21st century would attract the attention of the security
forces!
Overt observation
In an overt observation, participants
o are aware that they are being observed and may have been informed of
this in advance
Overt observations are more likely to occur in controlled lab conditions as the
researcher is keen to test the effect of the IV on the DV e.g.
As each of the above studies was a controlled observation it would not have
benefited the study to use covert methods
Limitations
Participants are aware that they are being observed and that their behaviour is
being measured which could give rise to participant reactivity
o The researcher may set up the observation schedule and tasks to align
too closely with their hypothesis
o Participants may not be aware that the researcher is an 'outsider' (in fact
it is highly likely that the observation is covert) e.g.
o In Piliavin's New York subway study, the observers noted that many of
the female passengers did not help in the emergency which could give rise
to further research on gender roles
Limitations
Participant observations could result in the researcher having a restricted
view of what they wish to observe and thus missing some important behaviours
o In Rosenhan's study, the researcher and confederates did not have full
access to every part of the hospital and all of the staff
o They may begin to identify with those they are observing, particularly
with long-term studies
Non-participant observation
In a non-participant observation:
o The researcher stays separate and apart from the group they are
observing
o Participants may or may not be aware that they are being observed
The researcher is more likely to have a good vantage point from which to
observe behaviour as they are not restricted to particular times, rooms, areas or
locations which could occur with a participant observation
o This increases the scope of the observation so that more data can be
gathered
Limitations
Being removed and at a distance from the 'action' means that a non-participant
observation may lack key detail and insight only made possible through the
use of participant observational methods
As the researcher is apart from what they are observing it is possible that they
could misinterpret some behaviours
o helping/not helping
o the number of times shoppers go over to look inside a box when the
instruction 'Look inside this box' is placed next to it
o Rather than recording every single behaviour available, they focus only on
the predetermined areas of interest
Evaluation of structured observation
Strengths
Using quantitative data is a quick and easy method which can be
presented visually in graphical form or converted to percentages and
statistics
o They can ignore any behaviours which do not align with the behavioural
categories they have decided upon
o This ensures that what is being observed is relevant to the research aim
Limitations
Quantitative data can shed light on what was observed but not on why that
behaviour occurred
Unstructured observation
An unstructured observation may be chosen by a researcher when
observing small samples in more intimate
environments where interpersonal interaction is the focus of the observation
o The data from the observation session(s) can then be used in conjunction
with other methods
Limitations
Due to the highly personal and subjective nature of unstructured
observations, the researcher may lose their sense of objectivity
o They may use confirmation bias when analysing their record of the
sessions
o This means that the published findings may not be a valid account of the
observation process
Behavioural categories
Behavioural categories are used to record specific behaviours during one
observation session
These categories could then be arranged into whether it is boys or girls who
are being aggressive/non-aggressive and to whom the aggression/non-aggression
is directed e.g.
Even when categories of behaviour are firmly established an observation can still
be affected by researcher bias
Researchers can test the reliability of their observations by comparing them with
another researcher's recording of their behaviours
Establishing good inter-observer reliability means that there is less chance that
researcher bias has interfered with the observation
o This means that subjectivity and the need to interpret the behaviour is
eliminated
The use of more than one observer should ensure inter-observer reliability
o Being able to claim reliability means that the research is less likely to
be criticised during the peer review process
Limitations
The predetermined behavioural categories may be limiting in terms of the
types of behaviours enacted during an observation session
o If one or more behaviours recur without there being categories for them
then this means that the research does not accurately represent what
occurred during the session
One issue with inter-observer reliability is that it does not allow for the possibility
that raters simply guessed rather than scoring the categories according to strict
criteria
o event sampling
o time sampling
o the frequency with which children in a classroom raise their hands to ask
a question
With time sampling the researcher records all behaviours during a set time
frame, at a set point e.g.
The researcher decides which time sample is most appropriate for that specific
piece of research
Time sampling allows the researcher flexibility to record any behaviours which
may be relevant to the research
Limitations
If too many of the specific behaviours occur at once and are overly complex it is
difficult for event sampling to capture them all
o This limits the validity of the method as it would not provide a true
reflection of what occurred during the observation session
Time sampling can miss any behaviours that occur outside of the set time frame
o This limits the validity of the method as some behaviours will be over-
represented in the findings
Correlations
To make sure whether a correlational relationship is causal, the two variables must be
investigated in a laboratory environment where extraneous variables are controlled.
Research Process
Null hypotheses: any relationship that is found between the variables is purely
due to chance.
Variables
Pilot studies are conducted to analyse the technical and financial risks and to assess
the feasibility of the study. Any plausible confounding variables are found and controlled
to ensure it does not affect the real trial.
Standardised procedures are important to ensure that all participants undergo the
same procedure. This helps to increase reliability and replicability.
Sampling participants
Random sampling: all participants are chosen randomly. Could be with a draw,
or random number generator.
Strengths – The sample is likely to be representative of the target population as
all type of people has an equal chance of being chosen.
Weakness – Everyone may not be equally chosen. For example, there could be
more girls chosen randomly than boys.
Qualitative Data: data written in a non-numerical format that often expresses a quality
or opinion.
Strengths: highly valid, unrequested, but important data is incurred.
Weaknesses: data interpretation may be subjective. Not representative, generalisable, or
reliable.
Data can be analysed using the measure of central tendency such as the mode, median,
and mean, and measures of spread such as the range, and standard deviation.
The measure of central tendency: a mathematical way to find the average score from
a data set using the mode, median, and mean.
Mean is calculated by adding all the scores in a data set and dividing them by the
number of scores in the data set.
Median is the middle score of a data set when it is ranked in order (ascending order)
The range is between the most significant and negligible values with an addition 1.
Standard deviation calculates the average difference between each score and the data
set's mean.
The normal distribution is an even spread of a symmetrical variable about the mean,
median and mode. It forms a bell-shaped curve and is symmetrical.
Bar charts are graphs used for data in discrete categories and total or average scores.
There are gaps between the columns as the data is not related linearly.
Ethical Considerations
Right to withdraw: they should be informed that they can withdraw at any
point.
Deception: participants should not be deceived during the study however if it’s
necessary to do so to protect the findings of the study, then participants need to
be debriefed.
For animals:
Species and strain: chosen species/strain should be least likely to suffer pain.
Whether they were socially housed or participated in other studies.
Housing: Isolation and crowding can cause animals distress. Caging conditions
should depend on the social behaviour of the species. Overcrowding ➔ distress &
aggression.
Evaluating Research
Reliability – the consistency of the outcome.
Validity – the extent to which the study measures what is intended to study.
Ecological Validity – the extent to which the results of the study represent real-
life behaviour.
Generalisability – the extent to which the results represent the behaviour of the
target population.
Test-retest: a way to measure the consistency of a test. The test is used twice
and if the scores on both tests are similar, then it has good reliability.
Validity
Validity is the extent to which the researcher is testing what they claim to be testing.
Internal validity is how well an experiment controls confounding variables. This allows
the researcher to be more condent about the causal relationship.
Ecological validity is the extent to which the findings in one situation would
generalise to other situations. This is influenced by whether the situation represents
the real world effectively and whether the task is relevant to real life.
Mundane realism is the extent to which a task represents the real-world situation.
Concurrent validity is when a test correlates well with a measure that has previously
been validated.
Demand characteristics are features of an experiment that give away the aims. This
could cause participants to change their behaviour and hence reduce the validity of
the study.
Reliability
Internal reliability refers to whether the procedures are standardised so that each
participant experiences the same thing.
External reliability is the extent to which the results of a procedure can be replicated
from one time to another, gaining consistent results.
There are two methods to test the reliability of a study: the split-half method and the
test-retest method
The split-half method involves the results of the first half of the questionnaire or
interview to be the same as the results of second half when the questions are the
same in both halves but presented in different a manner.
The test-retest method is a way to measure the consistency of a test or task by using
it twice and then comparing the results of each time to check how similar they are.