Social
Work
Statistic
s
Sr. Naomi Ruth O. Solatorio,
RSW
Objectives
To review the To To utilize
basic demonstrate learnings
concepts, learnings and and
operations, knowledge knowledge
and through in acing the
procedures of participation SWLE
Social Work in the
Review Contents
Basic
Concepts
Descriptiv
e
Statistics
Inferentia
l
Statistics
Statistics
Include numerical facts and figures
One out of 10 mothers suffers postpartum
depression
The latest earthquake in Mindanao measured
5.8 on the
Richter scale
65% of the residents living in Misamis
Occidental was
affected by the recent typhoon
Statistics
“Refers to a range of techniques
and procedures for analyzing,
interpreting, displaying, and
making decisions based on
data”
Research
Cycle
Research
Cycle
Review Contents
Basic
Concepts
Two Types of Statistics
Descriptiv
Inferential
e
Statistics
Statistics
Descriptive
used to organize and
Statistics summarize
typical numerical values within a
data set
numbers that are used to
summarize and describe data
they do not involve generalizing
beyond the data at hand
Inferential Statistics
utilize more complex procedures
and calculations to generalize and
draw conclusions about a
population based on a sample
from that population
Inferential Statistics
the process of drawing
inferences about the whole
(the population) from a
subset of it (the sample)
VARIABLES
• commonly represented by
symbols X, Y, Z
• it may contain numerical or
nonnumerical meanings
• anything the researcher can
quantify, measure, or
categorize about the
Example:
We are interested in studying the
school performance of social work
students in City G
VARIABLES
• the characteristics of the
population under study is called
variables if it can take on two or
more different values among
the population units
Example:
We are interested in studying the
school performance of social work
students in City G.
Example:
school, age, gender, year level,
family income = variables
VARIABLE VALUE VALUE CATEGORY
possible responses are not
responses to a scaled but
measurement categorical
- In quantitative research
- Ex: income – value
gender preference – value
Table 1. Ethnicity and Monthly
Participants Income
Ethnicity Monthly Income
1 Manobo P5,000
2 Mandaya P10,000
3 Bla’an P8,000
4 Bagobo P15,000
Table 1. Ethnicity and Monthly
Participants Income
Ethnicity Monthly Income
1 Manobo P5,000
2 Mandaya P10,000
3 Bla’an P8,000
4 Bagobo P15,000
DISCRETE CONTINUOU
VARIABLE S
take only VARIABLE
those whose
finite VARIABLE
values are interval
responses, or ratio
CLASSIFICATION
can be arranged
such as
on a number line
nominal or (scale) without
ordinal scale breaking or being
Course Income
values omitted
Number of children Scores obtained in an
exam
TYPES OF VARIABLES
DEPENDENT VARIABLE INDEPENDENT VARIABLE
Variable that can be May cause or have
affected by other an effect to the
variables dependent variable
Variable that is (can be)
manipulated by the
researchers
Example
Effectiveness of the four
types of antidepressant on
the relief of depression
Example
Effectiveness of the four
types of antidepressant on
the relief of depression
Independent
Variable
Example
Effectiveness of the four
types of antidepressant on
the relief of depression
Dependent Variable
QUANTITATIVE DATA QUALITATIVE DATA
things, objects,
characteristics, attitudes,
words or codes
behaviors, or personality representing a
traits collected from an category or a class of
observation that can be
counted or represented as an the sample
amount
TYPES OF DATA
QUANTITATIVE DATA QUALITATIVE DATA
use numbers to
represent
do not imply a
characteristics, numerical
events, attitudes, ordering
personality traits, and
so on
TYPES OF DATA
FOUR LEVELS OF
NOMINAL INTERVAL
MEASUREMENT
ORDINAL RATIO
one simply reveals the
names or rank order ordering of an interval
categorizes cases or values, does scale with the
responses values of the not have a additional
variable in zero point, property that
do not and indicates its zero position
terms of the indicates the
imply any degree given the exact
distance absence of the
ordering to any quantity being
among the characteristic between
measured.
responses them
NOMINAL
1. What is your gender?
_______ Male ________ Female ________
Others
2. Have you ever attended a seminar related
to social work?
a. Yes, I have.
b. No, I have not.
c. Decline to state
ORDINAL
1. How satisfied are you to the social services of
DSWD?
a. Not satisfied
b. Somewhat satisfied
c. Satisfied
d. Very satisfied
2. How confident are you in pursuing your
course?
a. Very confident
b. Slightly confident
INTERVAL
1. What is the monthly salary of a newly
employed social worker in Agency X?
_______ / month
2. What is the temperature at 8 am and 11
am?
3. What is the age of person A and person B in
June 2023?
RATIO
1. How many siblings do you have?
2. What was your annual household income last
2021-2022?
Review Contents
Descript
ive
Statistic
s
Table
Table of
of contents
contents
Frequency
Frequency 01
01 Distribution
Percentage 02
02 Valid Percentage
Cumulative Cumulative
Frequency 03
03 Percentage
Measures of Central Mean, Median,
Tendency 04
04 Mode
Range, Mean
Measures of Deviation,
Variability 05
05 Standard
Deviation
Frequency
Frequency
Frequency refers to the
number of times an event or
a value occurs.
symbol = f
Frequency
Frequency
Distribution
Distribution
a representation, either in a
graphical or tabular format,
that displays the number of
observations within a given
interval.
(Young, 2020).
Types
Types of
of Frequency
Frequency Distribution
Distribution
Type of Contra-
UNGROUPED FREQUENCY ceptive f
DISTRIBUTION Abstinence 14
It shows the frequency of an item
in each separate data value rather Condoms 47
than groups of data values.
Injectables 2
Ex: How many people use certain
type of contraception? Pill 35
Abstinence – 14, Condoms – 47,
Injectables – 2, Pill – 35, None - None 307
307
Total (n) 405
Types
Types of
of Frequency
Frequency Distribution
Distribution
Scores f
GROUPED FREQUENCY
DISTRIBUTION 0 – 10 15
In this type, the data is arranged
11 - 20 10
and separated into groups called
class intervals. 21 - 30 5
The frequency of data belonging Class
to each class interval is noted in a 31 - 40 Intervals
17
frequency distribution table.
The grouped frequency table 41 - 50 8
shows the distribution of
Total (n) 55
frequencies in class intervals.
Ungrouped
Ungrouped VS
VS Grouped
Grouped Data
Data
Ungrouped Data Grouped Data
Data that has been organized
Data that has not been organized into groups
into groups. Also called as raw
Ex. Scores
data
Scores f
1-2 1
Ex. Scores 3-4 3
2,5,6,8,4,5,6,3,2,1 5-6
7-8
4
2
Basic
Basic Terminology
Terminology ofof Frequency
Frequency
Distribution
Distribution
Lower Class Limit (Lower Limit) Scores f
smallest data value that can be included in the 0 – 10 15
class
Upper Class Limit (Upper Limit)
11 - 20 10
largest data value that can be included in
the class 21 - 30 5
Class Width (Class Size)
31 - 40 17
difference between the upper and
lower boundaries of any class
41 - 50 8
category
Total (n) 55
Class
Class Mark
Mark
the mid-value of a
given class interval
14 + 47
Types
Types of
of Frequency
Frequency Distribution
Distribution
Type of Con-
traceptive f cf
CUMULATIVE FREQUENCY
DISTRIBUTION Abstinence 14 14
It is the sum of the first frequency
Condoms 47 61
and all frequencies below it in a
frequency distribution. You have Injectables 2 63
to add a value with the next value
then add the sum with the next Pill 35 98
value again and so on till the last.
The last cumulative frequency will None 307 405
be the total sum of all
Total (n) 405
frequencies.
Percentage and
02 Valid Percentage
frequency (f)
% =total number of values (n)X 100
Formula of Percentage
Find the percentage of the people
Example: who are using the different types of
contraceptives.
Type of Contraceptive
f %
Abstinence 14 3.46
Condoms 47 11.60
Injectables 2 0.49
Pill 35 8.64
None 307
Total (n) 405
% = 14/405(100)
Find the percentage of the people
Example: who are using the different types of
contraceptives.
Type of Contraceptive
f %
Abstinence 14 3.46
Condoms 47 11.60
Injectables 2 0.49
Pill 35 8.64
None 252 62.22
Valid Total 350
Missing Data 55 13.58
Total (n) 405
% = 14/405(100) 99.99%
frequency (f)
Valid %= total number of
X 100
actual responses
Formula of Valid Percentage
Find the percentage of the people
Example: who are using the different types of
contraceptives.
Type of Contraceptive
f % Valid %
Abstinence 14 3.46 4
Condoms 47 11.60 13.43
Injectables 2 0.49 0.57
Pill 35 8.64 10
None 252 62.22 72
Valid Total 350 86.42
Missing Data 55 13.58
100%
Total (n) 405 V% =
14/350(100)
cumulative frequency (cf)
C% = total number of
X 100
observations (n)
Formula of Cumulative
Percentage
Example:
Type of Contraceptive
f cf c%
Abstinence 14 14 3.46
Condoms 47 61 15.06
Injectables 2 63 15.56
Pill 35 98 24.20
None 307 405 100%
Total (n) 405
C% =
14/405(100)
Activity
Activity
Suppose that 345 college students are
surveyed regarding whether they have traveled
abroad. Overall, the total number of persons
who had traveled abroad was 205. However,
after inputting every student’s response on a
spreadsheet, you noticed that 78 students did
not response to your question. What is the valid
percentage for this group of students who had
traveled abroad?
CATEGORICAL
CATEGORICAL DATA
DATA NUMERICAL
NUMERICAL DATA
DATA
data which is in the form data which is in the form of
of words rather than numbers.
There are two types of
numbers. numerical data.
Discrete – a count that
For example, involves integers (such as
colors, makes of cars, or frequency)
types of music, Continuous – an uncountable
number of values within a
subjects offered in BSSW range (such as height)
BAR
BAR GRAPH
GRAPH
It is used to display the frequency of
different categories in a clear and
concise manner. Each category is
represented by a bar, with the height
of the bar proportional to the
frequency or count of that category.
*X-axis (horizontal axis) - represents
the variable being measured
*Y-axis (vertical axis) - represents the
frequency count
HISTOGRAM
HISTOGRAM
It is used to display the distribution of
continuous data by dividing the data into
intervals or bins and plotting the frequency
of data points within each bin.
*X-axis (horizontal axis) - represents the
intervals or bins
*Y-axis (vertical axis) - represents the
frequency of data points within each bin
Time Intervals Frequency (Number of Stu-
dents)
4:00 – 4:30 20
4:31 – 5:00 30
5:01 – 5:30 40
5:31 – 6:00 50
6:01 – 6:30 40
6:31 – 7:00 30
7:01 – 7:30 20
PIE
PIE CHART
CHART
It is used to display the frequency of data
points as a proportion of the total
frequency. It is a circular chart divided into
slices, with each slice representing a
category and the size of the slice
proportional to the frequency or count of
that category.
LINE
LINE GRAPH
GRAPH
A line graph is a unique graph which
is commonly used in statistics. It
represents the change in a quantity
with respect to another quantity or
relationship between two or more
sets of quantities.
.
The point of interaction is the origin.
Usually, the time component is
plotted along the x-axis, while the
corresponding observation is plotted
along y-axis.
DOT
DOT PLOTS
PLOTS
BOX
BOX PLOTS
PLOTS
FREQUENCY
FREQUENCY POLYGON
POLYGON
QUALITATIVE
QUALITATIVE QUANTITATIVE
QUANTITATIVE
VARIABLES
VARIABLES VARIABLES
VARIABLES
Pie Chart Histogram
Bar Graph Line Graph
Dot Plots
Box Plots
Frequency Polygon
04
Measures of
Central
Tendency
Population Mean
Mean
Mean
• The mean is the sum of the value of
each observation in a dataset divided by
Sample Mean
the number of observations. This is also
(Ungrouped Data)
known as the arithmetic average.
• It is used if the most reliable measure is
desired and when there are few with
very highs and a few with very low
Sample Mean
values
(Grouped Data)
ADVANTAGE
ADVANTAGE LIMITATION
LIMITATION
The mean can be The mean
used for both cannot be
continuous and calculated for
discrete numeric categorical data,
data. as the values
cannot be
Summation of Frequency
Mean
Total number of elements
Mean
Mean of
of Ungrouped
Ungrouped
Type of Contra-
ceptive f Data
Data
Abstinence 14
Condoms 47
Injectables 2
NO
Pill 35
TheMEAN
mean cannot be
None 307 calculated for categorical
data, as the values cannot
be summed.
Mean
Mean of
of Ungrouped
Ungrouped
Data
Data
Final Grade
per Subject
85 83
82 80
85 + 82 + 90 + 83 + 80
90 88 +
= 88
5086
=
Frequency
Summation of
Mean Class mark or
Midpoint
Mean Total number of elements
Mean of
of Grouped
Grouped Data
Data
Median
Median
Median
(Ungrouped Data)
●It is the midpoint or
middlemost of a
distribution
Median
(Grouped Data)
ADVANTAGE
ADVANTAGE LIMITATION
LIMITATION
The median is less • The median
affected by outliers and
cannot be
skewed data than the
mean and is usually the identified for
preferred measure of categorical
central tendency when nominal data, as
the distribution is not it cannot be
symmetrical
logically ordered.
Total number of elements
Place in the
distribution
Median
Median
Median of
of Ungrouped
Ungrouped
Type of Contra-
ceptive f Data
Data
Abstinence 14
Condoms 47
Injectables 2
NO MEDIAN
Pill 35
The median cannot be
None 307 identified for categorical
nominal data, as it cannot
be logically ordered.
Median
Median of
of Ungrouped
Ungrouped
Data
Data
Final Grade
per Subject
85 83
82 80
80, 82, 83, 85, 90
90 = th = 6 th
n=5 5+12 2
= 3rd
Median
83of
Median of Ungrouped
+ 85 = 168
Ungrouped
167 / 2 = 84
Data
Data
Final Grade
per Subject
85 83
82 80
80, 82, 83, 85, 88, 90
90 88 = th = 7 th
6+12 2
=
Total number of
Lower boundary of the
elements
median class
(frequency)
Class size or
class width
Median (cw)
Cumulative
Frequency of the median frequency before
class the median class
Median
Median of
of Grouped
Grouped
Data
Data
Mode
Mode
• The mode is the most
commonly occurring
value in a distribution.
Mode
(Grouped Data)
Mode
Mode of
of Ungrouped
Ungrouped
Data
Data
Type of Contra-
ceptive f Condoms
Abstinence 14 Most commonly occurring
value in the distribution
Condoms 47
Injectables 2 The mode has an advantage over
the median and the mean as it
Pill 35
can be found for both numerical
and categorical (non-numerical)
data.
Type
Type of
of Modes
Modes
[Link] Mode – one mode
[Link] Mode – two modes
[Link] Mode – three modes
[Link] Mode – four or more modes
[Link] mode at all
Lower boundary of the Difference between the highest
modal class frequency and the frequency
below it
Class size (i)
Mode or class width
(cw)
Difference between the highest
frequency and the frequency
above it
Mode
Mode of
of Grouped
Grouped Data
Data
Outliers
Outliers
A statistical outlier is a
value that, when plotted
as Y for a given X, lies far
beyond the average for
the scores. The outlier
can be a recording error
or an actual value.
05
Measures of
Variation
Overview
Overview
Statisticians use summary
measures to describe the amount
of variability or spread in a set of
data. The most common measures
of variability are the range,
variance, and mean deviation.
Range
Range
the difference between
r = highest value – lowest value the largest and smallest
values in a set of values.
r = highest value – lowest value
r = 41
Range
Range of
of Grouped
Grouped Data
Data
Class Interval Frequency
Identify the largest upper limit
47 - 52 1 and subtract it to the smallest
41 - 46 3 lower limit
35 - 40 5
29 - 34 8 r = 52 – 5
23 - 28 8 r = 47
17 – 22 7
11 - 16 4
5 – 10 4
Mean
Mean Deviation
Deviation
Mean Deviation
(Ungrouped Data)
Mean Deviation
(Grouped Data)
Summation of Absolute
value Difference
between value
of x and mean
Mean
Deviation
Total number of elements
Mean
Mean Deviation
Deviation of
of
Ungrouped
Ungrouped Data
Data
Summation of Frequency Difference
between the
midpoint and
mean
Mean
Deviation
Summation of frequency
Mean
Mean Deviation
Deviation of
of
Grouped
Grouped Data
Data
Variance
Variance
Population
Variance
Sample Variance
(Ungrouped Data)
Sample Variance
(Grouped Data)
Value of Population
Summation of
x Mean
Population Squared
Variance
Total number of
population
Populaton
Populaton Variance
Variance
Summation of Value of Sample Mean
x
Sample Squared
Variance
Total number of
sample
Sample
Sample Variance
Variance
Summation of Midpoint
Mean
Variance
Total number of elements
Variance
Variance of
of Grouped
Grouped (frequency)
Data
Data
Standard
Standard Deviation
Deviation
√ Standard(or σ) is a
A standard deviation
measure ofDeviation
how dispersed the
data is (Ungrouped
in relation to Data)
the mean.
√
Low standard deviation means
data are clustered around the
mean, and high standard
Standard
deviation indicates data are
moreDeviation
spread out.
(Grouped Data)
Summation of Value of Sample Mean
x
Square
Root
√ Squared
Total number of
sample
Standard
Standard Deviation
Deviation Square root of the
(Ungrouped
(Ungrouped Data)
Data) variance
Summation of Frequency Difference between
the value of x and
mean
Square Root
√ Total number of elements
(frequency)
Standard
Standard Deviation
Deviation of
of Square root of Variance
Grouped
Grouped Data
Data
Review Contents
Inferent
ial
Statistic
s
Normal
Distribution,
Z-score, and
Probability
Normal Distribution
121
Explaining Normal Distribution
A normal distribution is a type of continuous
probability distribution in which most data
points cluster toward the middle of the
range, while the rest taper off symmetrically
toward either extreme.
122
Explaining Normal Distribution
It fits many human characteristics, such as height,
weight, speed etc. Many living things in nature,
such as trees, animals and insects have many
characteristics that are normally distributed.
shows that data near the mean are more frequent in
occurrence than data far from the mean
123
In graphical form, normal distribution
appears as a “bell curve”
124
Which of the following is a
normal distribution?
125
Factors / Parameters of the Graph of the
Distribution
MEAN STANDARD DEVIATION
Determines the location of
Determine the shape of
the center of the bell-
shaped curve the graph particularly its
width
Central highest value of the
curve
126
Characteristics of a Normal Distribution
1. Bell-shaped Curve
127
Characteristics of a Normal Distribution
2. Normal
distributions are
symmetric around
the mean
128
Characteristics of a Normal Distribution
3. Mean, median,
and mode are equal
or coincides at the
center
*unimodal*
129
Characteristics of a Normal Distribution
3. Mean, median, and mode are equal or
coincides at the center
130
Characteristics of a Normal Distribution
4. Area under the
normal curve is
equal to 1.0
131
Characteristics of a Normal Distribution
5. Denser in the
center, less dense in
the tails
132
A picture is worth a thousand words
133
134
Example:
❏ The score of social work students in their
Midterm Examination is normally distributed
with a mean of 35 and a standard deviation of 5.
❏ What percent of the scores are between
30 and 40?
❏ What scores fall within 95% of the
distribution?
135
What percent of
the scores are
between 30 and
40?
136
What scores fall
within 95% of the
distribution?
137
Standard Normal
Distribution
138
Explaining Standard Normal Distribution
is the simplest case of normal
distribution
Mean is 0, Standard Deviation is 1
Any normal distribution can be
standardized by converting its values
into z scores.
139
140
What is the difference
between
normal deviation and
standard normal deviation?
141
DIFFERENCE
STANDARD
NORMAL NORMAL
Can take on any value as its Mean and standard
mean and standard deviation is always fixed
deviation
142
Standardizing Normal Distribution
❏This allows you to easily calculate the
probability of certain values occurring in
your distribution, or to compare data sets
with different means and standard
deviations.
143
How do we standardize?
1. Convert the individual value into z-score
144
Example:
❏The score of social work students in
their Midterm Examination is normally
distributed with a mean of 35 and a
standard deviation of 5.
❏Find the z-score that corresponds to
a score of x=58
145
Example:
x = 58 z = 48 - 35
Mean = 35 5
SD = 5 z = 13 / 5
z = 2.6
146
-3 -2 -1 0 1 2 3
20 25 30 35 40 45 50
147
148
Probability
149
150
Example
Suppose a class has
200
people in it, and 30
are seniors. What is
the probability of
picking a senior if
picked randomly?
Example
P = 30 / 200
= .15 or 15%
Hypothesis
Null
TestingAlternative
Hypothesis Hypothesis
50% 70%
An educated
Hypothesis -
hunch or testing
procedure for
speculation deciding whether
the outcome of a
prediction study
intended to be (results for a
tested in a sample) supports a
research study particular
theory or practical
innovation (which
Hypothesis is thought to apply
to a population).
Types of Hypothesis
Null
Example:
Ho: There is no
Hypothesis
statement that in the significant
population there is difference
no difference
(or a difference opposite
between A and B
to that predicted)
between populations; That means: A =
B
Types of Hypothesis
Alternative
Hypothesis
statement in
Example:
hypothesis testing about H1: There is a
the predicted significant difference
relation between between A and B
populations (often a
prediction of a difference That means: A
between population A<B
means).
A>B
Direction of Hypothesis
Directional
Hypothesis
Example:
One-tailed hypothesis
H1: Social workers who
researchers have more caseloads
are certain of or able to are more stressed than
predict the direction that those who have less
the relationship of the
variable under
investigation will fall on
the normal curve.
Direction of Hypothesis
Non-Directional
Hypothesis
Example:
Two-tailed hypothesis H1: Social workers who
have more caseloads
researchers believe that are more or less
significant differences stressed than those who
do exist but are unsure have less
or unable to predict the
direction of the
relationship
Procedures in Hypothesis-Testing
Chi- Correlati
01 square 02 on
03 T-test 04 ANOVA
CORRELATIO
COMPARING NO TESTING
N
Start SOMETHING RELATIONSH
COEFFICIEN
? IP?
YES T
COMPARING
NO ONE-
SAMPLE
GROUPS?
T/Z TEST
YES
YOU HAVE 2 NO
ANOVA
GROUPS?
YES YES
INDEPENDE ARE THEY NOINDEPENDE
NT DEPENDENT NT
T-TEST ? T-TEST
Chi-
0 square
1
Chi-square
• Used to test hypothesis for variables whose values
are categorical (nominal)
• Test for Goodness of Fit – single nominal variable
• Test for Independence – two nominal variable with
several categories
• More on comparing an observed frequency
distribution to an expected frequency distribution
Chi-square Tests
Goodness of Fit Independence
whether the frequency used to test whether two
distribution categorical
of a categorical variable is variables are related to each
different from your expectations. other.
Cross-tabulation
Table used to analyze categorical data
COLUMN
ROWS
0
2
Correlat
ion
Correlation
• measures the degree of association or
relationship between two variables. It helps to
understand how changes in one variable
relate to changes in another
How to Calculate Pearson’s R
Correlation Coefficient
3.
2. Identify
1. Make the
a table Transcribe
values of xy, x 4. Solve for
squared, y
r
squared, place
in the
corresponding
column
0
3
T-Test
Hypothesis Testing
using
One Sample T-test
Df = n-1
Hypothesis Testing
using
Independent T-test
Df = (n1+n2) - 2
Hypothesis Testing
using
Independent Sample
T-test
Hypothesis Testing
using
Independent Sample
T-test Df = n-1
Hypothesis Testing
using
Independent Sample
T-test