Chapter 1 Research Methods Notes
Chapter 1 Research Methods Notes
Research
Methods
Chapter 1 — Memorisation Notes
1.4 Observations
1.5 Correlations
1.10Ethical Considerations
★ Whole-Chapter Summary
1.1 Experiments
Aim Hypothesis
The purpose of the study — what the researcher wants to find out or the A precise, testable, falsifiable prediction based on the aim — it must be
question they want answered. possible to prove it wrong.
Laboratory experiment Good control of variables → raises validity. Causal relationships Artificial situation → participants may behave unnaturally → lower ecological validity.
Artificial, controlled can be established, since only the IV should affect the DV. Participants often notice the set-up and guess the aim → demand characteristics.
setting; participants Standardised procedures → high reliability and easy to replicate Easier to gain consent, but knowing they're being tested is itself an ethical
usually know they're being exactly. consideration.
tested
Field experiment Natural setting → participants behave naturally → higher Much harder to control extraneous variables → lower reliability and harder to
Participants' everyday, ecological validity. If participants are unaware they're in a study, replicate exactly. Less certain the IV (rather than something else) caused the change
natural environment; IV is demand characteristics are reduced. in the DV. Often impossible to gain informed consent beforehand.
still manipulated by the
researcher
Experimental Designs
Independent Different participants are used in each condition No order effects (each person only does the task Participant variables — pre-existing differences between
measures (level of the IV) once). Fewer demand characteristics, since each the two groups of people (age, ability, mood) might explain
participant only sees one condition. the result instead of the IV.
Repeated The SAME participants take part in every Individual differences can't bias the comparison, Order effects (practice/fatigue) from doing the task more
measures condition since each person acts as their own baseline. than once. Participants see every condition, so more likely
Needs fewer participants overall. to work out the aim (demand characteristics).
Matched Different participants, but each person in one Reduces participant variables without causing Time-consuming and difficult to find good matches;
pairs condition is matched with someone similar in the order effects, since each person still only does matching is rarely perfect, so unmeasured differences can
other condition (identical twins are ideal) one condition. still exist between pairs.
Threats to Validity
WORKED EXAMPLE
A researcher compares people working on soft chairs vs hard chairs to see which group works harder. If, by accident, all the "soft chair" participants
happen to also be arts students (who might naturally work differently to science students), "subject studied" becomes a confounding participant variable —
we can no longer be sure it was the chairs, rather than the subject group, that caused any difference in work rate. The fix: use random allocation so both
subject groups are evenly spread across conditions.
Ethics in Experiments
✓ Lab experiments: consent is usually easy to obtain, but deception may be necessary so participants don't guess the aim and create demand
characteristics.
✓ Field experiments: consent is often impossible to gain in advance, since participants may not even know a study is happening — this also makes their
right to withdraw unclear, since you can't withdraw from something you're unaware of.
✓ Privacy is easier to protect in a lab, because the tasks are pre-planned; a field setting carries more risk of accidentally invading someone's personal
space.
✓ Confidentiality matters in both, but is especially at risk in field studies if participants could be identified by, say, their workplace or school.
1.1 RECAP
Hypotheses: alternative (directional / non-directional) vs null — all variables operationalised.
IV = what's changed; DV = what's measured. Lab = high control, low realism. Field = high realism, low control.
3 designs: independent measures (order effects avoided, participant variables risk), repeated measures (opposite trade-off), matched pairs (middle
ground, but hard to arrange).
Confounding variables = participant variables (about people) + situational variables (about environment) — both threaten validity if uncontrolled.
1.2 Self-Reports
Questionnaires
Easy to analyse — you can total scores and calculate averages. But limited in Rich, detailed, valid data — participants aren't forced into a box. But harder to score consistently:
depth: a participant's true feeling might not fit any option offered, which can lower different researchers may interpret the same answer differently, which is a lack of inter-rater
validity. reliability.
Interviews
Semi-structured interview
A mix: some fixed questions (so answers CAN be compared) plus the freedom to ask extra follow-up questions specific to that person — generally seen as the best
balance.
1.2 RECAP
Self-reports = questionnaires (written) or interviews (spoken).
Closed → quantitative, easy to analyse but can miss true feelings. Open → qualitative, rich but harder to score reliably.
3 interview types: structured, unstructured, semi-structured.
Watch for: social desirability bias, response rate, subjectivity/inter-rater reliability.
1.3 Case Studies
Strengths Weaknesses
Very high validity — the person is studied in real depth, in a genuine real-life Subjectivity — researchers often build a close relationship with the participant, which can bias
context. Triangulation (multiple methods) supports validity further. how they interpret the data, lowering validity.
Ethical risk — such personal, detailed questions can feel intrusive; hard to keep the person's
identity confidential — even initials may not be enough if they're well known.
Low reliability — usually only one researcher and one participant, so it's hard to be sure the
interpretation is objective; a different researcher might interpret it differently.
Very low generalisability — findings are specific to that one person and may not apply to
anyone else at all.
1.3 RECAP
Case study = deep dive on ONE case, multiple methods/sources (triangulation).
Very valid, very detailed — but low reliability & generalisability, and real ethical risk.
1.4 Observations
Choice 1 — Setting
Choice 2 — Structure
Behavioural categories
The specific, operationalised actions being recorded — they break a continuous stream of behaviour into distinct, observable events (never an "inferred" mental state
you can't actually see, like "feeling happy").
Inter-observer reliability
How consistently two or more observers record the same event — checked by comparing their independent recordings of the same footage/session.
Naturalistic Behaviour is true-to-life → high ecological validity No guarantee the target behaviour will even happen
Controlled Ensures the behaviour of interest actually occurs Less natural → may reduce ecological validity
Unstructured Captures unexpected/important behaviours not predicted in advance Hard to record everything accurately → can lower reliability
Structured More reliable — observer only focuses on a small set of categories Might miss behaviours that aren't on the list
Covert Higher validity — no demand characteristics, since participants don't know they're Ethical issue (no informed consent); practically harder to arrange
watched
Overt Easier and more ethical to arrange Participants may change behaviour, knowing they're watched → lowers
validity
1.4 RECAP
4 choices: naturalistic/controlled · structured/unstructured · participant/non-participant · overt/covert.
Covert + naturalistic = most valid, but least ethical/practical.
Structured = most reliable (fewer categories to track, clearer definitions).
1.5 Correlations
Correlation Co-variables
A research method that looks for a relationship between two measured variables The two measured variables being compared in a correlation.
(co-variables) — a change in one is related to a change in the other.
Causal relationship
When a change in one variable is directly RESPONSIBLE for (causes) a change in another — establishing this always requires an experiment, never a correlation
alone.
1.5 RECAP
Correlation = relationship between measured co-variables (positive/negative/none).
Big rule to repeat in every answer: correlation ≠ causation — a third variable might explain both.
1.6 Longitudinal Studies
Cross-sectional study
Compares people at DIFFERENT ages/stages by testing different groups of people at ONE single point in time — the contrasting method to longitudinal.
REAL EXAMPLE
Terman's "Life Cycle Study of Children with High Ability" began in 1922 and followed highly intelligent children for decades. It found that a high IQ doesn't
guarantee success in life — but did produce famous graduates, including psychologist Lee Cronbach. Participants were nicknamed "Termites," and
knowing they'd been labelled "gifted" may itself have shaped their behaviour and choices.
1.6 RECAP
Longitudinal = same people, tested repeatedly over time. Cross-sectional = different ages, tested once.
Key strength: rules out cohort/participant differences as an explanation for change.
Key weakness: sample attrition shrinks and biases the sample over time.
1.7 Variables: Definition, Manipulation & Control
Operationalisation
Turning a vague idea into something clearly defined and measurable (e.g. "young" becomes "under 20 years old"). Applies to the IV/DV in experiments, the co-
variables in correlations, and the behavioural categories in observations.
Standardisation
Making sure every participant has exactly the same experience — same
instructions, procedure, equipment — no matter which condition they're in.
EXAMPLE: OPERATIONALISING
"Hard vs soft chairs" isn't precise enough → better: "chairs with wooden/plastic seats" vs "chairs with padded seats." "Students working better" isn't
precise → better: "number of homework pieces handed in on time" or "minutes spent on extra work."
1.7 RECAP
Operationalise everything: IV, DV, co-variables, behavioural categories.
Confounds = participant variables (about people) + situational variables (about environment).
Pilot study catches problems early; standardisation keeps every participant's experience identical, raising reliability.
1.8 Sampling of Participants
Population Sample
The whole group of people who share a certain characteristic, from which a The smaller group actually selected to take part, ideally representative of the
sample is drawn. population.
Opportunity Choosing whoever is conveniently available Quick and easy — larger samples can be gathered fast Likely unrepresentative — available people
sampling tend to be alike (e.g. all similar
age/background)
Volunteer An advert/announcement is put out; those who Easy to arrange; volunteers tend to be committed (e.g. willing Volunteers often share traits (e.g. more free
sampling (self- respond become the sample to return for retesting); good for finding rare/unusual time, more curious) → unrepresentative
selected) participants
Random Every member of the population has an EQUAL Most likely to be genuinely representative Time-consuming to arrange properly; still
sampling chance of selection (e.g. numbers drawn from a biased if the original population list is
hat) incomplete
1.8 RECAP
3 techniques: opportunity (convenient), volunteer (self-selected), random (equal chance for everyone).
Random sampling tends to be most representative but is hardest to arrange in practice.
1.9 Data and Data Analysis
Types of Data
Strengths Usually objective; reliable scales; unusual but important responses aren't Often more valid — participants can express themselves fully instead of being
hidden by averaging; easy to compare using averages/spread. squeezed into fixed choices.
Weaknesses Collection method may limit what participants can express, lowering validity if More subjective — recording/interpretation may be biased by the researcher's own
their true view doesn't fit the options given. views; results from a few individuals may not generalise widely.
Mode The most frequently occurring score (can be more than one if tied) Any data, including categories Only average usable with categories, but ignores the actual
(e.g. favourite subject) values of scores → least informative
Median The middle value once scores are ranked smallest → largest (average Numerical/linear scale data only Unaffected by extreme outliers, but ignores the exact
the middle two if an even number of scores) values of most scores
Mean Sum of all scores ÷ number of scores Numerical/linear scale data only Most informative (uses every value), but CAN be
distorted/skewed by one or two extreme outliers
Spread (dispersion)
Graphs
Bar chart Data in separate/discrete categories (e.g. totals, means or modes for different Bars have GAPS between them — the categories aren't part of one continuous
groups) scale
Histogram Continuous data (e.g. a whole distribution of test scores) Bars TOUCH — the x-axis is one continuous scale (DV on x-axis, frequency on
y-axis)
Scatter Correlational data (two co-variables) Each dot = one participant's score on BOTH variables; a line of best fit may be
graph added
1.9 RECAP
Data: quantitative (numbers) vs qualitative (descriptions).
Averages: mode (most common), median (middle), mean (sum ÷ count).
Spread: range (quick, crude) vs standard deviation (uses every score, more accurate).
Graphs: bar chart (categories, gaps) · histogram (continuous, no gaps) · scatter graph (correlations).
1.10 Ethical Considerations
Human Participants
Deception Confidentiality
Avoid where possible; if truly necessary, plan to minimise distress and debrief Data stored securely and never released; names replaced with numbers/initials;
fully afterwards. institutions anonymised too.
Privacy Debriefing
Never invade a participant's physical/emotional space; they can decline to
answer; only observe where they'd expect to be seen anyway. Full explanation given after the study so participants leave in at least as positive
a state as when they arrived.
MUST KNOW
✓ Extra care is needed with children, people with mental health conditions/learning difficulties, non-native speakers, or groups who might feel pressured
(e.g. prisoners) — their capacity to give truly informed consent may be limited.
✓ Debriefing does NOT replace designing an ethical study in the first place — it can reduce harm afterwards, but it can't undo real damage that's already
been done.
Animal Research
1.10 RECAP
Humans: informed consent, right to withdraw, protection from harm, avoiding deception, confidentiality, privacy, debriefing.
Animals: Bateson's cube (benefit + quality vs suffering); replacement, appropriate species/numbers, humane procedures.
1.11 Evaluating Research: Methodological Issues
Reliability
ELABORATE
Reliability can be improved through: standardisation (same instructions/procedures/materials for everyone), clear operational definitions, and
researchers discussing/training together to interpret data consistently.
Validity
Replicability Generalisability
Whether a study's exact procedure can be repeated (by the same or different How widely a study's findings apply beyond the specific sample studied —
researchers) to check the results — requires very detailed, clear reporting of the depends heavily on how representative/large the sample was, and on ecological
original method. validity.
MUST KNOW: THE 4-QUESTION CHECKLIST (USE FOR ANY STUDY, INCLUDING UNFAMILIAR ONES)
✓ Is it valid? Does it test what it claims to? Think ecological validity, demand characteristics, subjectivity.
✓ Is it reliable? Are the measures/tools consistent? Could interpretation of the data be subjective?
✓ Is it generalisable? Would the findings apply to other people/places/times? Depends on the sample.
✓ Could it be replicated? Is there enough procedural detail reported to repeat it exactly?
1.11 RECAP
Reliability = consistency (test-retest, inter-rater, inter-observer).
Validity = measures what it claims (face validity, ecological validity).
Generalisability = applies elsewhere. Replicability = the study can be repeated exactly.
Use the 4-question checklist to evaluate ANY study — including unseen exam scenarios.
★ Whole-Chapter Summary
HYPOTHESES VARIABLES
Alternative: directional / non-directional IV / DV (experiments)
Null (chance explains any difference) Co-variables (correlations)
Confounds: participant + situational
All must be operationalised
METHODOLOGICAL EVALUATION
Reliability — test-retest, inter-rater, inter-observer
Validity — face, ecological
Generalisability & replicability
FINAL TIP
Cover each flashcard and try to recall it before checking the answer.
If you can explain WHY a rule exists (not just state it), you can apply it to unfamiliar exam scenarios too.