Basic Research Methods Notes
Basic Research Methods Notes
Take general claim, make it able to be examined empirically by making it more specific, take
groups, administer for example standard surveys and then analyse this data through proper
statistical methods
Depending on the hypothesis said and the examination of data it can be said that the
hypothesis is either supported or not supported
Empirical questions- Asking questions about the phenomenon that can be answered through
systematic empiricism
One of the most difficult aspects of psychological research is creating empirical questions that
are both interesting and empirically testable, as there are a lot of questions that simply can not
be answered empirically, at least for the time being.
Happiness- being in a positive mood all day, less depressive symptoms, having more serotonin
This makes it measurable. All concepts involved in an empirical question must be measurable.
MEASURABILITY is strictly related with technology, example of 500 years ago not being able to
prove whether dead people can think, and we now have mri scanners
If we for example find that good memories are often more recalled than bad memories we can
be tempted into conclude they are better remembered, however we cant rule out that for
example a person just has more happy memories or something of the sort. We need a reliable
MEASURE of the memories against people can be compared with (in memory of course)
Possible approaches (at least sum): Ask a parent to recollect positive and negative memories of
their childhood and compare this with the recollection of the person so that they may be
compared
2. initiate the study during childhood and ask parent to list diary of major events, question
individual as adult and compare
PUBLIC KNOWLEDGE: Science creates public knowledge, an ever-growing social process based
on the growth of knowledge across space and time
In regards to respect, psychology has an issue with respect as the other sciences that use the
empirical approach commonly examine phenomenon that are often not tangible and usually
not subjects of debate, which makes the average person hold them in a higher opinion while
psychology studies phenomenon closely intertwined with common sense and intuition
1 human reasoning is often driven not solely by logical thinking but often by content and
context specific aspects.
2 humans don’t tolerate uncertainty well, and as such they tend to classify things as true or
false, when often they can be untestable, uncertain or partially true
3 common sense and gut feeling is good enough for everyday situations (sometimes not even
those) but not for the complex functioning of the human mind
4 Human beings tend to look for results that support their prior belief, confirmation bias, which
makes it difficult to do revise knowledge and generate new knowledge
5 Empirical studies require adeptness of observing, memory and analysis that humans do not
naturally posses
6 tend to endorse opinions of what they deem to be experts
Scientists need
Skepticism- to stop during research and ponder possible alternatives and alternative
evidence, especially empirically collected
Scientific training- to conduct research with established and proper statistical methods,
and also consulting and knowing the scientific literature, taking years of training
High tolerance for uncertainty: sometimes it isn’t possible to prove or disprove a
hypothesis so we simply need to be able to say that its uncertain with the evidence
possessed at the moment
Research conducting
Design study (How will you collect data? From who? When where?)
Contact participitants
Collect data
Question must be interesting (unanswered, fills a gap in literature) and feasible (realistic
planning with regards to money time and everything else)
Professional journals –
main types of articles: metaanalysis, tutorials, empirical research report (one or more new
studies and results), review reports(report of the literature and can include propostition of new
approach or interpretation of literature and results gotten up till now)
Profecionel zurnels
Often assessed through impact factor, indicates number of citations obtained by each article
published in the journal, divided into quartiles usually , q1-q2 best
• IF is sensitive to how many people work in a discipline: a) Comparisons of IF make sense only
for a specific discipline b) A journal IF should also not be used to evaluate single publications (or
the persons writing them)
Still use this shit.
Report findings truthfully and honestly , stick to principles of open science as much as possible
Open science
Yappa yappa make available to everyone, non open access more prestigious usually and
researchers no hav mony for publish, still in non oa journals can publish to research gate and
allat
3 Seek justice: award adequate compensation for participant, example of distributing lijek
efektivni
4 respect ppl rights and dignity, informed consent important document, should include
description, duration procedures, possible discomfort, adverse effects, ability to withdraw once
started experiment, foreseeable consequences of this, prospective research benefits, limits of
confidentiality and incentive for participation
IN a lot of studies degree of deception is included, which is controversial but this can be
remedied through a debriefing upon finishing study.
Some amount of participant privacy violation is unavoidable, except anonymous surveys, but
try to keep as confidential as possible
Ethics committee, senior researchers at institution that give pass or fail and can suggest
possible improvements for ethical reasons ofc
3 lekcija
Phenomenon- well established fact that was replicated in empirical research several times,
neeed to be replicated or else either its wrong or needs context inspection to define
discrepancy leading to clarification
Theory in psychology
Coherent explanation of one or more phenomena, usually goes beyond observable phenomena
in question and has abstract concepts, variables and other things, importance of parsimony, use
only as many concept as u need to explain it, generate hypothesis, hypodeductive method
where a hypothesis is either confirmed or disconfirmed, leading to the hypothesis either
supporting or weakening a theory, issue of induction, cant be sure, white n black swan example,
and when it is not supported it leads to either theory revision or examination of the hypothesis
testing
Paradigm
Fayerabend. Examine through history why a certain theory was accepted or unaccepted, due to
often reasons outside of science, historical periods n shiet
Focault, practically same shit as Fayerabend, examined through episteme of the epoch
4 lekc
Aim of research is to get info on a reference population, but studies often conducted on
population sample. Sampling is the choice of the sample of population
Colpa di feasibilita
Generalizing, u know, golden rule is that a good constructed sample is similar in all important
aspects to the population so it can be generalized unto iT HUSSAH
Sampling error, results cant be generalized onto population because it is BURGER. Bad sample
Sample size recommendation? Many variables that may affect should have bigger sample, and
logically less variables smaller sample
Random methods:
Random
Completely random should in theory lead to generalizability but small sample can make this
fucked, mediate this by stratified sample which collects sample through different strata, where
strata are proportional to the population. These methods usually not used for practical reasons
Convenience sampling
Also theres the WEIRD acronym for usual population that is tested
5 lekc
Variables nooo
Levels of variables, each second one has properties of one before but not the one after it
Categorical, no quantity just a category, diff values called diff levels of the variable
1 group members that are the same in measured variable must be placed in the same variable
level
Ordinal, numbers don’t have metric meaning but are used to measure order of level or somn,
indicates whether someone has less or more of a variable
Ordinal must follow 2 principles aside from the 2 of the one before:
If one member has a larger amount of variable than other he is assigned a bigger number
Interval, quantity present but no natural zero, used to determine how much the different
individuals difer, need to know true ratio of differences for this and any numbers that
accurately express true ratio are acceptable
Linear transofmrations of an interval scale of measurement will also result in an interval scale of
measurement of that variable
N1- (a*n1)+b and so on for all numbers, where N1 is the transformed number, n1 is orginal
number, and a and are real numbers that a is not ZERO
Variable has meaningful 0 only when the 0 means the variable is absent
Limit of interval variables is that the number ratios aren’t meaningful, so one level isn’t that
many times bigger than this one etc
Keypoints of interval scales are that 0s are arbitraty and don’t represent complete lack of
variable
N1 is a x n1
Argument that psycho variables can at most be measured using ordinal scale, form a theoretical
perspective we should suspend judgement and wait for advancement but from a practical
standpoint we use mostly interval level scales for purpouses of statistical methods
Operational definition of a variable is how that variable can be measured, can be more than 1
operational definition
Lekc 6
Basics of JAAASP
Lekc 7 six seve
Descriptive statistics
Descriptive statistics is a set of methods used for summarizing and displaying data
Difference between frequency and measurement Is that measurement is assigning values to the
variable while the frequency is the number of people of a certain level of the variable in the
population sample
In frequency table the scores are assigned from largest to smallest, and although categorical
variables can be displayed through frequency tables and histograms the shape of the
distribution is not able to be displayed with them
Probability is defined as the p being equal to the number of observations of a level I divided by
total number of observations pi equal ni divided by ntot
Normal distribution often seen, unimodal and symmetrical distribution but an unimodal and
symmetrical distribution is not always normal distribution
Most statistical models used in psychological research assume that the distribution is normal
The tests for normality are as follows, visual representation, statistical and numeric test
Histogram can be used in jasp along with the density line, with interval option setting the
intervals in the variable that is measured.
Numerical test, in jasp we can also quantitatively see the normality by selecting statistics and
then skewness coefficient
Skewness coefficient larger than 1 is positive, 0 is normal and smaller than -1 is negative.
Between -1 and 1 is considered a normal distribution
Statistical test
Shapiro Wilk normality test measures whether the distribution of the variable differs in a
statistically significant manner from the normal distribution. This is inferential statistics where
two hypothesis are displayed, n0 and n1 with n0 being the null hypothesis and n1 being the
alternative hypothesis
Data inferences made by reference population of sample, which are used to either discard null
and accept alternative or vice versa
P value, depends on context, here is probability of obtaining the observed distribution in our
sample , assuming the variable follows normal distribution in the reference population from
where it was inferred from
If it isn’t it is h1
This test has its problems, with when the sample is small, rule of thumb less than 20 it can be
inaccurate and also when it is subject to sampling bias or a sampling error in general so it
doesn’t represent the population accurately
Central tendency of a distribution is where the most common values are distributed in the
variable, we have the most common 3 measures of it, mean mode and median
Mean is everything summed up divided by number of elements, mode is the most common
appearing value and median is the value at the middle
Mode is the only measure of central tendency that can be applied to categorical scales of
measurement
Variables measured on ordinal have median and mode but not mean, while ratio n interval
have it all
Mean is by far the most commonly used in ratio and interval scales but it is highly subject to
being affected by the distribution, while if the case of the distribution is symmetrical then the
mode mean and median are relatively the same
In case of strongly skewed distributions the median is preferred as it is not affected as the mean
is
Since the c t measures depend on dispersion measures like the mean should be used with at
least one measure of dispersion
Dispersion deals with the variability of the data within the distribution
Three most important measures of distribution are standard deviation, interquartile range and
and range. Can only be applied to interval or ordinal scales of measurement
Simplest and least informative is range and it is the maximum and minimum value difference in
the distribution
The interquartile range describes the difference between the third quartile and the first
quartile, where the 3rd quartile is where below 75% of values fall under and 1st quartile is below
25% fall under. The median quartile is 50%, the 2nd
Percentile rank of a score measures which values fall under that %. We can calculate this by
organizing scores from lowest to largest, count how many of them are lower than it, divide
number by total and multiply by 100
In raw terms standard deviation is the mean distance between the values In a distribution and
the mean of the distribution
Divide xi where in the variable x an I score is there, take away from it the mean score of the
variable x, sve to u zagradi, kvadriraj zagradu, dodaj sve zajedno I podijeli s brojem ljudi u
uzorku
We also have variance which is standard deviation squared, same info as SD. Both should be
expressed with the units that correspond to them
Standard error is the standard deviation divided by the square root of the sample numerosity
Wtf
Outliers- problem where there is a sample member whose scores differ significantly from other
scores
This can fuck the data up, not obious by standard deviation or other numerical representations
of dispersion, cause the data dispersion might be relatively small but these outliers can skew
them. Best way to spot them is a histogram, through visual analysis
Outliers can be a significant problem because they strongly interfer with the mean, in some
extreme cases making it a lie!
What do we do with them? No rule, for example when we are sure that they result from
systematic bias, we can remove them, this is reasonable.
Example of this is if for example duhhhhhhhhh we are sure that the participant did not
understand the questions
Rule of thumb is to ensure in general that the main conclusions of a study are not substantially
dependent on the presence or lack of outliers
We may conduct the analysis both with and without them and if the results are robus they will
remain consistent regardless of their presence and absence.
We can describe the relative position of a score through the z score where from the score the
mean is subtracted and then the result of that is divided by the standard deviation of the scores
This is called the standardization of scores and the z score is called a standardized score
We should when missing data occurs see the potential reasons for thei absence such as the
question being too complicated or overly intrusive. We can deal with this 2 way, 1 ignore, 2
replace wit mean 😊
Independent and dependent value with independent being the one that affects and dependent
the one being affected
We consider cases only involving quantitative, usually interval measurement scale as they are
most common for psychology
In graphs the independent variable is usually expressed on the horizontal axis and the
dependent on the vertical axis
Bar graphs, usually used for categorical variables or quantitative variables with a small number
of levels (2-10, usual rule of thumb)
Line graphs, equivalent to bar graphs, representing the mean n stnd err values for the
dependend variable at diff levels of the independent variable
In jasp the process for creating and computing graphs depends on the type of research design,
for now focus one 1 ind 1 dep variable
First, we see if ind variable has 2 levels or more than 2 levels. If 2 then t test if more then anova
We must also see if it is a between subject design or a within subject design with a between
subject design being that the participant experiences only one level of the independent variable
while the within subject designs all participants experience all levels of the independent
variable (within n between designs most common for one ind variable research)
Between subject design example is group as ind variable where group has levels like patient
control or yip yap and yop and each participant is in one of these values
Within subject design example is time where each participant is tested at one interval then the
next, then additional times if designed so
When we check for p to prove h0 we must note that it is only acceptable measure if the both
levels of independent variable are normally distributed which is easily checked in jasp
Bar n line graph limitations are that they provide no dispersion data and no visualization of
outliers
Raincloud plots work with each of the cases mentioned up till now and seem very niiice.. :)
Scatterplots applicable to between subject designs where the independent variable is
quantitative and has many levels
Two common situation in which two diff types of effect size indices must be used
Situation 2: Independent quantitative with many possible values and dependent quantitative
Does bio sex affect in? compare mean scores, but we gotta check for dat dispersion. First case
std deviation difference is 2 something but its small, we see it visually and its aight and then 2 nd
case big standard deviation, like 13. This is a no no ☹
This example shows that for any numerical index of difference between groups the dispersion
must be taken into account for. Most common option for this is cohens d where the mean
scores that are subtracted are divided by the standard deviations of both the groups together.
• A Cohen’s d of 0.50 means that the two group means differ by half a pooled standard
deviation. A Cohen’s d of 1.20 means that they differ by 1.20 pooled standard deviations.
For the cases where we want effect size and the ind variable has more than 2 levels then the
equivalent of Cohens DICK is partial eta squared
Tends to have a strong positive value when ind variable score higher than mean of x score
higher on dep variable mean of y and same for score low
Lecture 8
When there is a statistical relationship found between variables this could mean that there is a
causal relaitionship between them as in the ind variable causes the shift in the dep variable, but
this does not HAVE to be the case.
Directionality problem
The choice of which variable influences which (ind and dep) is usually based on the researchers
hypothesis which is usually based on a definite theory
In some cases the directionality problem can remain partially or totally unresolved for reasons
of ethical nature.
In other cases the problem doesn’t exist because switching the roles of the variables makes no
sense, example of iq causing differences in bio sex
A confounding variable problem is that there exists a third variable that partially or completely
explains the relationship between the independent and dependent variable
In the case of bio sex influencing driving skills the confounding variable of driving experience
would be called a mediator, it mediates because It doesn’t affect biological sex, while the
atmospheric pressure listed above would be labeled a redundant relationship and a redundant
confounding variable
The answer to this again is experimenting, specifically where the experimenter tries to control
as many confounding variables as they can.
Lecture 9
Experimental research
Experimetns r useful for testing the causal relationship between variables. This is crucial for
testing hypothesis and theories
The experimenter controls extraneous variables (any variables not ind or dep)
Manipulate and control have specific meanings in experimenting, as manipulating means
systematically changing the levels of a variable in the experiment while controlling a variable
means preventing it from changing systematically in the experiment
The manipulation of the ind variable is meant to solve the directionality problem
The point of controlling extraneous variables is that they solve the confounding variable
problem.
If we control the extraneous variables, no change in dep variable, then manipulate the ind
variable and the dep variable changes we can say that the changes of the dep variable are due
to the effects of the ind variable
Possible levels of the manipulated independent variable are also called the conditions of the
experiment
In some cases the independent variable can only be manipulated actively in an indirect manner.
Example of mood. This is because the manipulation could be mediated by other factors
In these cases a manipulation check is conducted where the independent variable is tested
before and after the manipulation attempt
In some cases for practical or ethical reasons the independent variable cannot be manipulated,
ex of child illness. This doesn’t mean that the relationship isn’t able to be studied, it just means
it must be done through non experimental designs
There is usually a large number of extraneous variables, for example the comfort of the
participant seems to be an important situational variable, however the situational extraneous
variables usually constitute a minor problem for researchers as they do not typically change
systematically during the experiment
Observable and measurable extraneous variables can also be controlled very easily
The process of sitributing the effect of a measurable extraneous variable is called the
counterbalancing of that variable
However this may lead to an impediment of the generalizability of the experiment results
Some of these variables are unmeasurable though, and they can constitute a bigger problem,
how do we solve this? Ramdagadam dagadigadigadam
Randomization implies the random assignment of the participant to the conditions of the
experiment, and if the sample is large enough this ensures that the distribution of personal
extraneous variables is approximately the same across all the conditions
Exp design p2
Two general measures of the overall quality of the psychological study are internal validity and
external validity
Internal validity
A study considered high in internal validity means that the design ensures that the change is the
dependent variable is CAUSED by the independent variable’s manipulation
The highest level of this can be achieved when it is concluded that confounding variables had
no effect on the dependent variable
External validity
A study considered high in external validity is one whose results can be generalized to people
outside of the sample of the experiment and to contexts outside of the one the study was
conducted in. For psychology this is very important as we seek to understand general laws of
human psychological functioniong
As a general rule the study has high external validity when the sample that the researcher
chose closely resembles the population onto which the researcher wishes to generalize the
results on
Regarding situations studied, psycho exp are often criticized by the need for controlling of
extraneous variables and the manipulation of the independent variable, causing the conditions
to appear artificial, not as they are irl.
In many cases there are key diffs between this artificial setting and real situations
Computers and other measurement devices are used in the majority of psychological
experiments, they allow for great control of extraneous variables, the implemention of
standardized, automated procedures for data collecting, and to be very precise in measuring
the patients response.
They increase internal validity, however they reduce external validity as there is a difference
between how people behave in front of a computer and in real situations
Between subject designs are widely used in clinical research. Example of taking sample of
agoraphobic people and assigning them randomly to one of 3 treatment options
Internal validity of this type of research can be high provided the experimenter controls for
extraneous variables effectively
Often used to determine whether a treatment, intervention or training works
Often there is a group that receives the treatment, and then the control group which doesn’t.
this is called a randomized clinical trial, it is experimentu… -_-
The basic logic behind it is that simply seeing a difference before and after in the dep variable
isn’t sufficient to declare the treatment effective
We should do an effect size estimation on deez.
Lets say the treatment is effective, control group gets it. Lets say its not, both groups get
alternative treatment. Simplifies ethical concerns.
To avoid difference of ppl in treatment group getting placebo cause they know they getting
treated and the control group not getting it, we introduce the placebo control group. Self
explanatory
Another option is removing the control condition entirely and replacing it with th best available
alternative treatment. For one placebo reduced, secondly if an effective alternative exists, its
only important whether this new one is better or not
Primary disadvantage carryover effect, being tested in one condition may affect the behaviour
while being tested in the next condition
At least 3 types
Practice effect – participants perform better in next conditions cause they had a chance to
practice it. Effect can be particularly strong in difficult tasks or response time tasks
An artifact is a result that is not caused by the phenomenon under study but by the arbitrary
choices of the experimenter. If not controlled for the practice effect could become an artifact.
Type of variable that an artifact could be is a situational confounding variable
Fatigue effect- opposite of the practice effect, may perform task worse due to fatigue or
boredom, especially significant in long and or fatiguing tasks. Especially in studies where
response time is measured and in which the number of trials is very large
Context effect- Effect in which the previous conditions influence the response of the successive
conditions through their experiences. For example if a chud were to be showed beforehand,
the perceived attractiveness of the current normie would be increased, while it may decrease if
a chad was shown beforehand. This type of context effect is also called the contrast effect
Stroop effect
Sometimes between better such as we cant test treatment with withinsubject, and cant ask
them to recall memories as they already recalled them in a positive mood, in a negative mood
they will most likely recall the same memories
Refers to any systematic variation of the dependent variable which is due to the experimenters
expectancy to how participants should behave.
This effect decreases both internal and external validity of the experiment
Hawthorne effect, related to the effect where peoples behaviour improves as a reaction to the
fact that they are being tested
This also causes problems for external and internal validity of the experiment
John Henry effect- individuals who are aware that are at a disadvantage might perform with
more effort in an effort to close the gap between them and the other group
They are called constructs because they are not obvious or selfevident, but they are
«constructed» upon a conventional conceptual definition.
Constructs are unobservable both because they are mostly internal but also because they refer
to tendencies
The operational definition of a construct is how the construct should be measured. Three
categories
Self report measures- report own thoughts, feelings, actions. Widely used in clinical psychology
Physiological measures: recording the activity on a physiological level ofone or more of the
participants body parts
Large number of constructs can be measured in multiple ways, and sometimes they are as this
can potentially deepen our understanding of the relationship of the ind and dep variable (s in
this case)
In one experiment the experimenter may measure more than one construct. Example of the
professors experiment of expectancy and perceived weight
Factorial designs
The most common approach to measuring multiple variables is a factorial design, where every
level of one independent value is combined with each level of the other indepent variable/s to
produce all possible combinations. Also the ind variables are called factors
Each cell represents one condition, with each condition benig a combination of the levels of the
ind variable (factors)
Number of ind variables and number of levels of these variables can be represented by
multiplying each no of levels in the first factor to the n factor
In mixed factorial designs at least one factor is between and one is within subject
A main effect refers to the individual influence of a specific factor on the dependent variable
In a factorial line plot, non-parallel lines indicate an interaction between the factor on the
horizontal axis and the factor represented by the different lines.
In factorial designs, interaction effects are often more interesting and more informative than
main effects.
Read ts again
Don’t manipulate ind var, don’t cnt extr var, not experimental
Example is epidemiological study, which estimate prevalence of a specific disease within a
population
CORRELATIONAL RESEARCH
Distinction between ind and dep variable becomes much less relevant
Example 2 The relationship between the two measured variables is thought to be causal, but
researchers cannot manipulate the independent variable because it is impossible, impractical,
or unethical.
which studies the effects of brain damage on cognitive functions, such as thought, emotions,
and behavior.
Some third variables that may affect the dependent variable can be controlled to some extent
using complex correlational designs, where potential confounding variables are treated as
independent variables (e.g., “the influence of the extent of parietal lobe damage AND drug use
on logical skills”).
The relationship between brain damage and cognitive functions can only be studied by means
of correlational research. • However, the more general relationship between brain and
behaviour can also be studied using experimental designs…
Naturalistic observation is when the observer seeks to collect data I nthe environment of the
behaviour in which it occurs
It’s a type of correlational research because it lacks manipulation of the independent variable
Involves manipulation of independent variable but does not involve random putting individuals
in conditions
4 types
Non equivalent group designs are between subject designs where the individuals are not
randomly assigned to conditions, which is the step that allows researchers to prevent
systematic personal extraneous variable interference
In clinical research another possible omission of the results is spontaneous remission where
physiological and psychological issues can get better over time by themselves
A treatment is judged to be effective only when the patients in the treatment condition
improve more than the patients in the control condition (no treatment and/or placebo control
conditions).
Pretest psottest designs widely used in educational research to test effectiveness of new
education programs
Sometimes ppl that only differ significantly to the mean are selected but sometimes this is not
concrete as regression to mean is possible
Interrupted time series designs are a variant of pretest-posttest designs
• Interrupted time-series designs are a variant of pretest-posttest designs. A time series is a set
of measurements taken at intervals over a period of time. In an interrupted time series-design,
a time series is interrupted by a treatment. • For example, suppose that the dependent variable
is the number of aggressive acts taken by a group of inmates. • Suppose that the DV is
measured once a month for 6 months. Then, after 6 months, the inmates get a 3-month
psychotherapy for the reduction of aggressive acts. The number of aggressive acts is then
measured once a month for other 3 months.
Because history, maturation, and regression to the mean are likely to be similar for the students
in both schools, any observed difference between the improvements of the two groups of
students is probably due to the effects of treatment. • However, this type of design does not
completely prevent the possibility of confounding variables, for two reasons: 1) Something
could occur at one of the schools but not the other (e.g., a student drug overdose), so students
at the first school would be affected by it while students at the other school would not. In other
words, there might be different “histories” for the two schools. 2) Similar to the case of
nonequivalent groups designs, the treatment group and the control group may differ in some
important extraneous variables that the researchers could not control.
The problem of internal validity in quasi-experimental research: a summary The internal validity
of quasi-experimental designs should be evaluated “case-by-case”. Researchers can enhance
the internal validity of quasiexperimental research designs by controlling factors that may
potentially lower internal validity, such as spontaneous remission, history, and maturation.