0% found this document useful (0 votes)
4 views54 pages

Statistics

The document provides an overview of statistics, including definitions, types of data, and methods for representing and analyzing data. It covers real-life applications of statistics, measures of central tendency, and various graphical representations such as histograms and frequency polygons. Additionally, it discusses the calculation of mean, median, mode, and quartiles for grouped data.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
4 views54 pages

Statistics

The document provides an overview of statistics, including definitions, types of data, and methods for representing and analyzing data. It covers real-life applications of statistics, measures of central tendency, and various graphical representations such as histograms and frequency polygons. Additionally, it discusses the calculation of mean, median, mode, and quartiles for grouped data.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

LET'S LEARN

STATISTICS
P U T R I R A Z N I A S A F I R A , M . P d .
S M A P R O G R E S I F B U M I S H A L A W A T
ISLAMIC VALUE
ISLAMIC VALUE

‫ب ََل يُغَا ِد ُر‬


ِ ٰ
‫ت‬ ‫ك‬
ِ ْ
‫ال‬ ‫ا‬ َ
‫ذ‬ ‫ه‬ٰ ‫ل‬
ِ َ‫ا‬ ‫م‬ ‫َا‬ ‫ن‬َ ‫ت‬َ ‫ل‬ ‫ي‬
ْ ‫و‬
َ ٰ
‫ي‬ ‫ن‬
َ ‫و‬ ُ ‫ل‬ ‫و‬ُ
ْ ْ َ َ ‫ق‬‫ي‬ ‫و‬ ‫ه‬
ِ ‫ي‬
ْ ‫ف‬
ِ ‫ا‬‫م‬َّ ‫م‬
ِ ‫ْن‬
َ ‫ي‬ ‫ق‬
ِ ‫ف‬
ِ ْ
‫ش‬ ‫م‬
ُ ‫ْن‬َ ‫ي‬ ‫م‬
ِ ‫ر‬‫ج‬ْ
ِ ُ ‫م‬ ْ
‫ال‬ ‫ى‬ ‫ر‬
َ َ ‫ت‬َ ‫ف‬ ‫ب‬
ُ ‫ت‬ٰ ‫ك‬
ِ ْ
‫ال‬ ‫ض َع‬ ِ ‫َو ُو‬
‫۝‬٤٩ ‫ك ا َ َحدًا‬ ْ َ‫اض ًر ۗا َو ََل ي‬
َ ُّ‫ظ ِل ُم َرب‬ ِ َ ‫ح‬ ‫ا‬ ‫و‬
ْ ُ ‫ل‬ ‫م‬
ِ ‫ع‬ َ ‫ا‬ ‫م‬
َ ‫ا‬ ‫ُو‬ْ ‫د‬ ‫ج‬ ‫و‬ ‫و‬ ۚ
َ َ َ ‫َل ا َ ْحصٰ ى َه‬
‫ا‬ ٓ َّ ِ‫ص ِغي َْرة ً َّو ََل َك ِبي َْرة ً ا‬
َ

"Put a book (record of deeds) on each person, then you will see sinners feeling
afraid of what is (written) in it. They said, "How wretched we are, what book is
this, leaving nothing out of the small and the great, except to record them." They
found (all) what they had done (written). Your Lord does not wrong anyone.".
(Q.S. Al-Kahfi, 18: 49).
Real Life Applications
01. Weather Forecasting

02. Medical Studies

03. Sales Tracking

04. Quality Testing

05. Traffic Analysis


MINDMAP
01. Introduction

02. Representations of Data

03. Measures of Central Tendency

04. Measures of Position

05. Measures of Spread


Introduction
Definition of Statisics and Data
The collection, organisation and analysis of numerical information are all part of the subject called
statistics.

Pieces of numerical and other information are called data.

What are Types of Data in Statistics?


Qualitative or Categorical Data
describes the data that fits into the categories. Qualitative data are not numerical. The categorical information involves
categorical variables that describe the features such as a person’s gender, home town etc. Categorical measures are
defined in terms of natural language specifications, but not in terms of numbers.

Quantitative or Numerical Data


Quantitative data is also known as numerical data which represents the numerical value (i.e., how much, how often, how
many). Numerical data gives information about the quantities of a specific thing. Some examples of numerical data are
height, length, size, weight, and so on.
Representations of Data
Why is representation of
data important in statistics?
Representations of Data
Benefits of Representation of Data
Grouped Frequency Distributions
When the range of the data is large, the data must be grouped into classes that are more
than one unit in width, in what is called a grouped frequency distribution.

Constructing a Grouped Frequency Distribution


1. Find the Range (R)
𝑹𝒂𝒏𝒈𝒆(𝑹) = 𝒉𝒊𝒈𝒉𝒆𝒔𝒕 𝒗𝒂𝒍𝒖𝒆 − 𝒍𝒐𝒘𝒆𝒔𝒕 𝒗𝒂𝒍𝒖𝒆
2. Find the number of classes (K)
𝑲 = 𝟏 + 𝟑, 𝟑 𝒍𝒐𝒈 𝒏
Where : n is the number of data
3. Find the class width (p) by dividing the range by the number of classes and rounding up.

𝑹
𝒑=
𝑲
Grouped FrequencyHistogram
Distributions
Solution
Step 1. Find the range
These data represent the record high temperatures in
degrees Fahrenheit (°F) for each of the 50 states.
Construct a grouped frequency distribution for the data.
Step 2. Find the number of classes (K).
112 110 107 116 120 100 118
Step 3. Find the class width. 112 108 113 127 117 114 110
120 120 116 115 121 117 134
118 118 113 105 118 122 117
Class Limit Tally Frequency 120 110 105 114 118 119 118
110 114 122 111 112 109 105
106 104 114 112 109 110 111
114
Source: The World Almanac and Book of Facts.
Total
Grouped FrequencyHistogram
Distributions
Solution
These data represent the record high temperatures in
degrees Fahrenheit (°F) for each of the 50 states.
Construct a grouped frequency distribution for the data.

112 110 107 116 120 100 118


112 108 113 127 117 114 110
120 120 116 115 121 117 134
118 118 113 105 118 122 117
120 110 105 114 118 119 118
110 114 122 111 112 109 105
106 104 114 112 109 110 111
114
Source: The World Almanac and Book of Facts.
Grouped Frequency Distributions

several things should be noted

Lower Boundary
= lower limit – 0.5
Lower Class Limit

Upper Boundary
Upper Class Limit
= upper limit + 0.5
Class Width
= upper boundary – lower boundary
The histogram is a graph that
displays the data by using
contiguous vertical bars
(unless the frequency of a
class is 0) of various heights
to represent the frequencies
Histogram of the classes.
Histogram
Record High Temperatures

Solution

Step 1. Draw and label the x and y axes. The x axis is


always the horizontal axis, and the y axis is
always the vertical axis.
Step 2. Represent the frequency on the y axis and the
class boundaries on the x axis.
Step 3. Using the frequencies as the heights, draw
vertical bars for each class.
The frequency polygon is a
graph that displays the data
by using lines that connect
points plotted for the

The Frequency frequencies at the midpoints


of the classes. The

Polygon frequencies are represented


by the heights of the points.
The Frequency Polygon
Record High Temperatures

Solution

Step 1. Find the midpoints of each class. Recall that


midpoints are found by adding the upper
and lower boundaries and dividing by 2:

𝟗𝟗. 𝟓 + 𝟏𝟎𝟒. 𝟓
= 𝟏𝟎𝟐
𝟐
dst.
The Frequency Polygon
Solution

Step 2. Draw the x and y axes. Label the x axis with the midpoint of each class, and then use a suitable
scale on the y axis for the frequencies.
Step 3. Using the midpoints for the x values and the frequencies as the y values, plot the points.
Step 4. Connect adjacent points with line segments. Draw a line back to the x axis at the beginning and
end of the graph, at the same distance that the previous and next midpoints would be located.
This type of graph is called the
cumulative frequency graph, or
ogive.
The ogive is a graph that
represents the cumulative
frequencies for the classes in a

The Ogive
frequency distribution.

The ogives are also of two types:


1. Less than Ogive
2. More than Ogive
Record High Temperatures
Ogive

Solution (Less than ogive)


Step 1. Find the cumulative frequency for each class..
Ogive
Solution

Step 2. Draw the x and y axes. Label the x axis with the class boundaries. Use an appropriate scale for the y axis to
represent the cumulative frequencies.
Step 3. Plot the cumulative frequency at each upper class boundary.
Step 4. Starting with the first upper class boundary, 104.5, connect adjacent points with line segments. Then extend the
graph to the first lower class boundary, 99.5, on the x axis.
Let’s Contruct
More than ogive
Let’s Try!
Protein Grams in Fast Food The amount of protein (in grams) for a
variety of fast food sandwiches is re ported here. Construct
a frequency distribution. Draw a histogram, a frequency polygon,
and an ogive for the data.

23 30 20 27 44 26 35 20 29 29 25
15 18 27 19 22 12 26 34 15 27 35
26 43 35 14 24 12 23 31 40 35 38
57 22 42 24 21 27 33

Source: The Doctor’s Pocket Calorie, Fat, and Carbohydrate Counter.


Measures of
Central Tendency
Do you still remember the mean, median, and
mode that you learned in junior high school?
Recall
Mean, Median, and Mode for Ungrouped Data

The data show the number of public libraries in a sample of eight states. Find the
mean, median, and mode.
114 77 21 101 311 77 159 382
Source: The World Almanac.

Solution:
Histogram
Mean for Grouped Data
Definition

Mean is the average of the given numbers and Hence, mean is calculated using the
following formula :
is calculated by dividing the sum of given
numbers by the total number of numbers.
σ 𝑓. 𝑥𝑚
𝑥ҧ =
You can calculate the mean, the class containing
𝑛
the median, and the modal class for continuous
data presented in a grouped frequency table by 𝒇. 𝒙𝒎 is the sum of the product of the frequency
finding the midpoint of each class interval. (𝒇) and the midpoint (𝒙𝒎 ) for each class.
𝒏 is the total number of observations.
Histogram
Mean for Grouped Data
Example
The frequency distribution shows the salaries (in Make a table as shown.
millions) for a specific year of the top 25 CEOs in
Class Frequency Midpoint
the United States. Find the mean. Boundaries (𝒇) (𝒙𝒎 )
𝒇. 𝒙𝒎
Source: S & P Capital.

𝒏= ෍ 𝑓. 𝑥𝑚 =

Find the mean:

Recall that midpoints are found by adding the


upper and lower boundaries and dividing by 2
Histogram
Median for Grouped Data
Definition
Hence, median is calculated using the following
formula :
The median is the middle value when the data
𝑛
values are put in order. − 𝑐𝑓𝑏
𝑀𝑒 = 𝑙 + 2 ×𝑐
𝑓
To find the median class, we have to find the Where;
𝒏
cumulative frequencies of all the classes and . 𝑙 is the lower boundary of the median class
𝟐

After that, locate the class whose cumulative 𝑛 is the total number of observations
𝒏 𝑐𝑓𝑏 is cumulative frequency of the class before the
frequency is greater than (nearest to) . The
𝟐
median class
class is called the median class.
𝑓 is the frequency of the median class

𝑐 is the class width


Histogram
Median for Grouped Data
Example
The frequency distribution shows the salaries (in
millions) for a specific year of the top 25 CEOs in
Solution
the United States. Find the median.
Source: S & P Capital.
Histogram
Mode for Grouped Data
Definition
The mode or modal class is the value or class
that occurs most often. Example

Hence, mode is calculated using the following The frequency distribution shows the salaries
formula : (in millions) for a specific year of the top 25
CEOs in the United States. Find the median.
𝑑1 Source: S & P Capital.
𝑀𝑜 = 𝑙 + ×𝑐
𝑑1 + 𝑑2

Where;

𝑙 is the lower boundary of the modal class

𝑑1 is frequency of the modal class – frequency of the class before

𝑑2 is frequency of the modal class – frequency of the class after

𝑐 is the class width


Let’s Try!

Find the Mean, Median, and Mode.


Measures of Position
Histogram
Quartiles for Grouped Data
Definition
Quartiles divide the distribution into four equal
Step 3: Compute for 𝑸𝟏 , 𝑸𝟐 , 𝑸𝟑 . The
groups, denoted by 𝑸𝟏 , 𝑸𝟐 , 𝑸𝟑 . following formula is used
𝑖. 𝑛
− 𝑐𝑓𝑏
4
𝑄𝑖 = 𝑙 + ×𝑐
Where;
𝑓
Step 1: Determine the cumulative frequency 𝑖 is the value of quartile being asked
Step 2: Determine the 𝑸𝟏 , 𝑸𝟐 , 𝑸𝟑 classes
𝒏 𝑙 is the lower boundary of the 𝑄𝑖 class
The 𝑸𝟏 class is the class interval where the th data
𝟒
is contained. 𝑛 is the total number of observations
𝟐𝒏
The 𝑸𝟐 class is the class interval where the th data 𝑐𝑓𝑏 is cumulative frequency of the class before the
𝟒
is contained. 𝑄𝑖 class
𝟑𝒏
The 𝑸𝟑 class is the class interval where the th data 𝑓 is the frequency of the 𝑄𝑖 class
𝟒
is contained.
𝑐 is the class width
Histogram
Quartiles for Grouped Data
Example
The frequency distribution shows the salaries (in
millions) for a specific year of the top 25 CEOs in
Solution
the United States. Find the 𝑄1 , Q3 .
Source: S & P Capital.
The interquartile range (IQR) is the difference
between the third and first quartiles.

𝑰𝑸𝑹 = 𝑸𝟑 − 𝑸𝟏
Let’s Try!

Find the 𝑸𝟏 , 𝑸𝟐 , 𝑸𝟑 and IQR.


Deciles For Grouped Data
Definition
Deciles divide the distribution into 10 groups, as shown. They are denoted by 𝑫𝟏 , 𝑫𝟐 , etc.

Step 1: Determine the cumulative frequency


Step 2: Determine the decile classes Where;
𝒊 .𝒏 𝑖 is the value of decile being asked
The 𝑫𝒊 class is the class interval where the th
𝟏𝟎
𝑙 is the lower boundary of the 𝐷𝑖 class
data is contained.
𝑛 is the total number of observations
Step 3: Use the following formula to solve for 𝑫𝒊 𝑐𝑓𝑏 is cumulative frequency of the class before the
𝐷𝑖 class
𝑖. 𝑛
− 𝑐𝑓𝑏 𝑓 is the frequency of the 𝐷𝑖 class
10 𝑐 is the class width
𝐷𝑖 = 𝑙 + ×𝑐
𝑓
Histogram
Deciles for Grouped Data
Example
The frequency distribution shows the salaries (in
millions) for a specific year of the top 25 CEOs in
Solution
the United States. Find the 𝐷4 , D7 .
Source: S & P Capital.
Histogram
Percentiles for Grouped Data
Definition
Percentiles divide the data set into 100 equal Step 3: Use the following formula to solve
groups. Percentiles are symbolized by for 𝑷𝒊

𝑷𝟏 , 𝑷𝟐 , 𝑷𝟑 , . . . , 𝑷𝟗𝟗 . 𝑖. 𝑛
− 𝑐𝑓𝑏
100
𝑃𝑖 = 𝑙 + ×𝑐
Where;
𝑓
𝑖 is the value of percentile being asked
Step 1: Determine the cumulative frequency 𝑙 is the lower boundary of the 𝑃𝑖 class

Step 2: Determine the percentile classes 𝑛 is the total number of observations


𝑐𝑓𝑏 is cumulative frequency of the class before the
The 𝑷𝒊 class is the class interval where the
𝑃𝑖 class
𝒊 .𝒏
th data is contained. 𝑓 is the frequency of the 𝑃𝑖 class
𝟏𝟎𝟎
𝑐 is the class width
Histogram
Percentiles for Grouped Data
Example
The frequency distribution shows the salaries (in
millions) for a specific year of the top 25 CEOs in
Solution
the United States. Find the 𝑃55 , P60 .
Source: S & P Capital.
Any Questions?
Measures of Spread
Histogram
Standard Deviation for Ungrouped Data
Definition
Hence, standard deviation (𝝈) is calculated
using the following formula :
Standard Deviation is defined as the degree of
dispersion of the data points from the mean
σ𝒏𝒊=𝟏 𝒙𝒊 − 𝒙
ഥ 𝟐
value of the data points. 𝝈=
𝒏

A higher standard deviation means the data Where;

points are more spread out, while a lower 𝑥𝑖 is the 𝑖 𝑡ℎ observation

standard deviation means they are closer to the 𝑥ҧ is mean

mean. 𝑛 is number of observations

Σ is sum over all the values


Histogram
Standard Deviation for Ungrouped Data
Example
Find the standard deviation for the number
Solution
of months brand A lasted before fading was

10, 60, 50, 30, 40, 20


Histogram
Variance for Ungrouped Data
Definition
Hence, variance (𝝈𝟐 ) is calculated using the
following formula :
Variance for ungrouped data is the average of
the squared differences between each data 𝒏
σ𝒊=𝟏 𝟐
𝟐

𝒙𝒊 − 𝒙
point and the data set's mean. 𝝈 =
𝒏
It measures how spread out a set of data is, Where;

with a smaller variance indicating that data 𝑥𝑖 is the 𝑖 𝑡ℎ observation

points are close to the mean, and a larger 𝑥ҧ is mean

variance indicating they are more spread out. 𝑛 is number of observations

Σ is sum over all the values


Histogram
Variance for Ungrouped Data
Example
Find the variance for the number of
Solution
months brand A lasted before fading was

10, 60, 50, 30, 40, 20


Histogram
Mean Deviation for Ungrouped Data
Definition
Hence, mean deviation (𝑴𝑫) is calculated using
the following formula :
The mean deviation (also known as Mean
𝒏
σ𝒊=𝟏
Absolute Deviation, or MAD) of the data set is ഥ
𝒙𝒊 − 𝒙
the value that tells us how far each data point is 𝑴𝑫 =
𝒏
from the center point of the data set. The
center point of the data set can be the Mean, Where;

Median, or Mode. Thus, the mean of the 𝑥𝑖 is the 𝑖 𝑡ℎ observation

deviation is the average of the absolute 𝑥ҧ is mean

deviations of all data points from the chosen 𝑛 is number of observations

central value. Σ is sum over all the values


Histogram
Mean Deviation for Ungrouped Data
Example
Find the mean deviation for the number of
Solution
months brand A lasted before fading was

10, 60, 50, 30, 40, 20


Let’s Practice!
Combined Mean
Recall
Formula

Example 1:

60 students of section A of Class XI,

obtained 40 mean marks in statistics,

40 students of section B obtained 35

mean marks in statistics. Find out

mean marks in Statistics for class XI

as a whole.
Recall
Formula

Example 2:

The average score of 10 students is

80. If the scores of two new students

are added, the new average score

becomes 82. What is the average

score of the two new students?


Let’s Make Summary!
THANK YOU
S e e Y o u

You might also like