Understanding Data Management in Statistics
Understanding Data Management in Statistics
MODULE 8
DATA MANAGEMENT
STATISTICS is a scientific body of knowledge that deals with the collection, organization or presentation,
analysis, and interpretation of data.
• Collection refers to the gathering of information or data.
• Organization or presentation involves summarizing data or information in textual, graphical, or tabular
forms.
• Analysis involves describing the data by using statistical methods and procedures.
• Interpretation refers to the process of making conclusions based on the analyzed data.
BRANCHES OF STATISTICS
▪ Descriptive Statistics – is a statistical procedure concerned with describing the characteristics and
properties of a group of persons, places or things.
For example, we may describe a collection of persons by stating how many are poor and how
many are rich, how many are literate and how many are illiterate, how many fall into various categories
of age, height, civil status, IQ, and many more. We may also describe a particular barangay in terms of
the number of families it has, the number of grade-schoolers, the number of professionals, the number
of households with certain kinds of appliances, the number of siblings in each household, or the rate of
unemployment.
Generally, descriptive statistics involve gathering, organizing, presenting and describing data.
▪ Inferential Statistics – is a statistical procedure that is used to draw inferences or information about the
properties or characteristics by a large group of people, places, or things or the basis of the information
obtained from a small portion of a large group.
Suppose we want to know the most favorite brand of toothpaste of a certain barangay and we do
not have enough time and money to interview all the residents of that barangay; we may just ask selected
residents. With the data obtained from the interviews, we shall draw or make conclusions as to
barangay’s favorite brand of toothpaste. This example involves the use of inferential statistics.
TERMINOLOGIES IN STATISTICS
Some important terms are commonly used in the study of Statistics. These terms should be understood fully in
order to facilitate the study of statistics.
1. Population refers to a large collection of objects, places, or things. To illustrate this, suppose a
researcher wants to determine the average income of the residents of a certain barangay, and there are
1500 residents in the barangay. Then all of these residents comprise the population. A population is
usually denoted or represented by N. Hence, this case, N = 1500.
GE 4- Mathematics in the Modern World La Carlota City College
2. Sample is a small portion or part of a population. It could also be defined as a sub-group, subset, or
representative of a population. For instance, suppose the above-mentioned researcher does not have
enough time and money to conduct the study using the whole population, and he wants to use only 200
residents. These 200 residents comprise the sample. A sample is usually denoted by n, thus n = 200.
3. Parameter is any numerical or nominal characteristic of a population. It is a value or measurement
obtained from a population. It is usually referred to as the actual value. If, in the preceding illustration,
the researcher uses the whole population (N=1500), then the average income obtained is called a
parameter.
4. Statistic is an estimate of a parameter. It is a value or measurement obtained from the sample. If the
researcher in the preceding illustration makes use of the sample (n=200), then the average income
obtained is called a statistic.
5. Data (singular form is datum) are facts, or a set of information or observation under study. More
specifically, data are gathered by the researcher from a population or from a sample. Data may be
classified into two categories: qualitative or quantitative
a. Qualitative data are data that can assume values that manifest the concepts of attributes. These are
sometimes called categorical data. Data falling in this category cannot be subjected to meaningful
arithmetic. They cannot be added, subtracted, or divided. Gender and nationality are qualitative
data.
Gender is a qualitative dichotomous variable since an individual may take one of the two
values “male or female”. In an opinion poll, the response of an individual towards an issue, whether
to “go” for it, “against” it, or “undecided,” is an example of a qualitative trichotomous variable.
Smoking habits of an individual in different situations may be classified as “Always/Very Often”,
“Often”, “Seldom”, “Very Seldom”, or “Never”. This set of qualitative values is called a
multinomous variable.
b. Quantitative Data are data that are numerical in nature. These are data obtained from counting or
measuring. In addition, meaningful arithmetic operations can be done with this type of data. Test
scores and height are quantitative data.
6. A variable is a characteristic or property of a population or sample that makes the members different
from each other. If a class consists of boys and girls, then gender is a variable in this class. Height is
also a variable because different people have different heights. Variables may be classified based on
whether they are discrete or continuous and whether they are dependent or independent.
a. Discrete Variable
A discrete variable can assume a finite number of values. In other words, it can assume specific
values only. The values of a discrete variable are obtained through the process of counting. The number
of students in a class is a discrete variable. If there are 40 students in a class, it cannot be reported that
there are 40.2 students or 40.5 students, because a fractional part of a student can't be in the class.
GE 4- Mathematics in the Modern World La Carlota City College
b. Continuous Variable
A continuous variable can assume infinite values within a specified interval. The values of a
continuous variable are obtained through measurement. Height is a continuous variable. If one reports
that the height of a building is 15 m, it is also possible that another person reports that the height of the
same building is 15.1m or 15.12m, depending on the precision of the measuring device used. In other
words, the height of the building can assume several values.
c. Dependent Variable
A dependent variable is a variable that is affected or influenced by another variable.
d. Independent Variable
An independent variable affects or influences the dependent variable. To illustrate Independent
and dependent variables, consider the problem entitled, The Effect of Computer-Assisted Instruction
on the Students’ Achievement in Mathematics. Here, the independent variable is the computer-
assisted instruction, while the dependent variable is the achievement of students in mathematics.
7. Constant refers to the fundamental quantities that do not change in value; fixed costs and acceleration
due to gravity are examples of such.
SCALES OF MEASUREMENT
1. Nominal Scale – This is the most primitive level of measurement. The nominal level of
measurement is used when we want to distinguish one object from another for identification
purposes. In this level, we can only say that one object is different from another, but the amount of
difference between them cannot be determined. We cannot tell that one is better or worse than the
other. Gender, nationality, and civil status are of a nominal scale.
2. Ordinal scale – In the ordinal level of measurement, data are arranged in some specified order or
rank. When objects are measured at this level, we can say that one is better or greater than the other.
But we cannot tell how much more or how much less of the characteristic one object has than the
other. The ranking of contestants in a beauty contest, or siblings in the family, or of honor students
in the class is of an ordinal scale.
3. Interval Scale – If data are measured at the interval level, we can say not only that one object is
greater or less than another, but we can also specify the amount of difference. The scores in an
examination are on an interval scale of measurement. To illustrate, suppose Kensly Kyle got 50 in
a Math examination while Kwenn Anne got 40. We can say Kensly Kyle got a higher score than
Kwenn Ann by 10 points.
4. Ratio Scale – The ratio level of measurement is like the interval level. The only difference is that
the ratio level always starts from an absolute or true zero point. In addition, in the ratio level, there
is always the presence of units of measure. If data are measured in this level, we can say that one
object is so many times as large or as small as the other. For example, suppose Mrs. Reyes weighs
GE 4- Mathematics in the Modern World La Carlota City College
50 kg, while her daughter weighs 25 kg. We can say that Mrs. Reyes is twice as heavy as her
daughter. Thus, weight is an example of data measured in a ratio.
▪ Ungrouped data are data that are either not organized, or if arranged, could only be from highest to
lowest or lowest to highest.
▪ Grouped data - are data that are organized and arranged into different classes or categories.
Arranging the scores from the lowest to highest will facilitate the enumeration of important characteristics of
the data. The test scores of the 50 students in Calculus arranged from lowest to highest are shown below:
3 13 17 20 27 30 32 35 40 43
9 13 18 21 28 30 33 36 40 46
10 14 18 25 28 31 34 37 40 48
10 15 19 26 28 31 35 38 41 50
12 16 20 26 29 32 35 39 42 50
The highest scores obtained is 50 and the lowest is 3. Ten students got a score of 40 and above, while only 4
got ten and below. Generally, the students performed well in the test with 33 students or 66% getting a score of
25 and above.
B. Stem – and – leaf plot which sorts data according to a certain pattern. It involves separating a number
into two parts. In a two-digit number, the stem consists of the first digit, and the leaf consists of the
second digit. While in the three digit number, the stem consists of the first two digits, and the leaf
consists of the last digit. In a one-digit number, the stem is zero.
Table 1.1
Stem-and-leaf Plot of an arranged Test Scores in Calculus of 50 Students
Stem Leaves
0 3,9
1 0,0,2,3,3,4,5,6,7,8,8,9
2 0,0,1,5,6,6,7,8,8,8,9
GE 4- Mathematics in the Modern World La Carlota City College
3 0,0,1,1,2,2,3,4,5,5,5,6,7,8,9
4 0,0,0,1,2,3,6,8
5 0,0
By looking at the stem-and –leaf plot, we can easily rank the data or put them in order. Thus, the ten
lowest scores are 3,9,10,10,12,13,13,14,15 and 16 while the ten highest scores are
40,40,40,41,42,43,46,48,50 and 50.
C. Tabular- this form of presentation is better than textual form because it provides numerical facts in a
more concise and systematic manner. Statistical tables are constructed to facilitate the
analysis of relationships. Each class/subclass is assigned to a particular row or column and
figures for various classifications are noted in appropriate cells.
Advantages of Tabular Presentation
1. It is brief, it reduces the matter to the minimum.
2. It provides the reader a good grasp of the meaning of the quantitative relationship indicated in the
report.
3. It tells the whole story without the necessity of mixing textual matter with figures.
4. The systematic arrangement of columns and rows makes them easily read and readily understood.
5. The column and rows make comparison easier.
FREQUENCY DISTRIBUTION
A frequency distribution table is a table that shows the data arranged into different classes and the number
of cases that fall into each class.
3. CLASS BOUNDARIES or REAL or EXACT CLASS LIMITS are the numbers used to separate
class but without gaps created by class limits. The number to be added or subtracted is half the difference
between the upper limit of one class and the lower limit of the preceding class.
Example
Class interval Class boundaries
L.L – U.L L.C.B – U.C.B
16 – 20 15.5 - 20.5
21 – 25 20.5 - 25.5
26 - 30 25.5 - 30.5
4. CLASS MARKS are the midpoints of the classes. They can be formed by adding the lower and upper
limits and then dividing by 2.
Example:
Class interval class mark/midpoint (X)
16-20 18
21-25 23
26-30 28
2. Decide on the number of class intervals. There should not be too many to avoid many empty classes,
and there should not be too few to avoid long details. Use the formula suggested by Sturge.
k = 1 + 3.3 log N
3. Divide the range R by the number of class intervals (k) to obtain the size of the class interval:
i= R/ k or c = R / k
4. Starting from the larger integer less than or equal to the minimum score, construct class intervals of
size i until the maximum score is reached.
5. Set up the class boundaries.
6. Tally the scores in appropriate classes and then add tallies for each class to obtain the frequency.
7. Solve the class mark or midpoint of each class. This is obtained by adding the lowest class limit and
the upper class limit, then dividing by 2.
25 33 48 39 34 29
29 37 39 35 41 29
23 32 48 28 45 19
Range = 57 – 18 = 39
Table 1.3
Grouped Frequency Distribution for the Entrance Examination Scores of 60 Students
Class Limits Class Boundaries Tally Frequency Classmark
18-20 17.5 - 20.5 111 3 19
21-23 20.5 - 23.5 111 3 22
24-26 23.5 - 26.5 1111-1 6 25
27-29 26.5 - 29.5 1111 5 28
30-32 29.5 - 32.5 1111-11 7 31
33-35 32.5 - 35.5 1111-1111 10 34
GE 4- Mathematics in the Modern World La Carlota City College
Relative Frequency
Relative Frequency= 𝑇𝑜𝑡𝑎𝑙 𝑛𝑢𝑚𝑏𝑒𝑟 𝑜𝑓 𝑜𝑏𝑠𝑒𝑟𝑣𝑎𝑡𝑖𝑜𝑛
D. Graphical Presentation – this form is the most effective means of organizing and presenting statistical
data because the important relationships are brought out more clearly and creatively in virtually solid
and colorful figures.
Measures of Central Tendency are numerical descriptive measures which indicate or locate the center of the
distribution or data set.
MEAN
The mean of the set of values or measurements is the sum of all the measurements divided by the
number of measurements in the set.
Example 1. Below are the travel times in minutes spent by Kenneth in going to school last week.
Day Time Spent in Travelling
Monday 60 min.
Tuesday 45 min.
Wednesday 50 min.
Thursday 53 min.
Friday 47 min.
Σx 60 + 45 + 50 + 53 + 47
̅=
𝒙 = ̅=
𝒙 = 51 minutes
𝑛 5
Example 2. Find the mean of the following scores 85, 79, 90, 92, 80, 78, 75,93
Example 1. Below are Dona’s subjects and the corresponding number of units and grades she got for the first
grading period. Compute her grade point average. (GPA)
Subject Unit Grade
Math 1 80
English 1 82
Filipino 1 83
Science 2 81
Social Studies 1 80
GE 4- Mathematics in the Modern World La Carlota City College
PEHM 1.5 85
Technology & 2 82
HE
Therefore, Dona’s has the GPA of 81.95 for the first grading period.
Example 2.
GWA = 48/24=2.0
CLASSMARK FORMULA
ΣfX
̅ =
𝒙 where, x =measurement or score, f=frequency, X=class mark, n= total frequency
𝑛
ΣfX 1828
X= = = 45.7
𝑛 40
CODED FORMULA:
Σfd
̅ = Xam +( ) i where, Xam = assumed mean, f=frequency, d=coded deviation, N=total frequency, i=class
𝒙 𝑛
size
Class F X d d fd fd
Interval
16-23 1 19.5 -3 0 -3 0
24-31 3 27.5 -2 1 -6 3
32-39 6 35.5 -1 2 -6 12
40-47 12 43.5 0 3 0 36
48-55 10 51.5 1 4 10 40
56-63 8 59.5 2 5 16 40
N=40 Σfd= 11 Σfd= 131
If we assign 0 in the class limits with the highest frequency
Σfd 11
̅ = Xam +(
𝒙 )i , = 43.5+ (40) 8 = 43.5 + 2.2 = 45.7
𝑛
MEDIAN
Median is the middle value of a given set of measurements, provided that the values or measurements
are arranged in an array. An array is an arrangement of values in increasing or decreasing order.
1. In an English test, eight students obtained the following scores: 10, 15, 12, 18, 16, 20, 12, 14. Find
the median.
𝑁
− <𝑐𝑓𝑏
Median = lb + ( 2
)i
𝑓
where lb = lower class boundary of the median class
n = total frequency
<cfb = less than cumulative frequency before the median class
i = size of the class interval
f = frequency of the median class
Let us illustrate how to compute the median of grouped data using the distribution of the test scores of
40 students in Mathematics given in the example of the preceding section.
Example 1:
MODE
Mode is the value that occurs most frequently in a set of measurements or values.
A distribution may have only one mode. In this case, the distribution is said to be unimodal. Data that
have two values for the mode are said to be bimodal. It is also possible that the set of data is multimodal if
there are more than two values for the mode. If all the scores in a set of data occur only once, then the set of
data has no mode.
Example 1: The data on the number of times 10 mothers go to market every week are shown below.
Mother A B C D E F G H I J
GE 4- Mathematics in the Modern World La Carlota City College
Solution: The mode is 3. This means that the majority of the mothers go to market three times a week.
fm− fb
Mode = lbmo + ( 2𝑓𝑚−𝑓𝑎−𝑓𝑏 ) i
It is important to note that the formula for the mode given above holds only for unimodal distribution. For
multimodal distribution, the rough mode is given by the formula
Mode = 3(Median) – 2(Mean)
Let us use the distribution of scores of 40 students in Mathematics to illustrate how to compute the mode for
grouped data.
Example: Find the mode of the data whose frequency distribution is given below.
Class Interval f
16-23 1
24-31 3
32-39 6
40-47modal class 12
48-55 10
56-63 8
n=40
Notice that the class intervals are arranged from lowest to highest group. The modal class is the class
interval 40-47. The lower class boundary of the modal class is 39.5, the frequency of the modal class is 12, the
frequency below the modal class is 6, the frequency above the modal class is 10, and the size of the class
interval is 8. Substituting these values in the formula, we have
fm− fb
Mode = lbmo + (2𝑓𝑚−𝑓𝑎−𝑓𝑏 ) i
GE 4- Mathematics in the Modern World La Carlota City College
= 39.5 + ( 12- 6 )8
2(12)-10-6
6
= 39.5 +( 8)8
= 39.5 + 6
= 45.5 or 46
In case a distribution has at least 2 modes, a rough mode can be computed as follows:
Moderough = 3 Mdn – 2Mn
where Mdn = median
Mn = mean
Recall that the median is the value where 50% of the distribution fails or lies above it while 50% of the
distribution lies below it. The midpoint of the line segment is the median.
In other words, the median is the value that divides the distribution into two equal parts. We define the
quartiles, deciles, and percentiles similarly. These descriptive measures – quartiles, deciles, and percentiles
– are called fractiles.
▪ Quartiles are values that divide the distribution into four equal parts.
▪ Deciles are values that divide the distribution into ten equal parts.
▪ Percentiles are values that divide the distribution into 100 equal parts.
The first quartile (Q1) is the value where 25% of the distribution lies below it, while 75% of the
distribution lies above it. The third quartile (Q3) is the value where 75% of the distribution lies below it
while 25% of the distribution lies above it. Observe that the median is equal to the second quartile (Q2).
Step 3. Locate the score corresponding to the obtained position in the distribution starting from the
lowest score.
Step 4. Interpolate to get the score if the obtained position from step 2 is not exact.
𝑛+1 𝑛+1 𝑛+1
Formula: Px = x( 100 ) Dx = x( 10 ) Qx = x( )
4
𝑥𝑛
−<𝑐𝑓𝑏
Qx = lb + ( 4 𝑓
)𝑖
Find the Q1, P36, & D8 of the data whose frequency distribution is given below.
1. Q1 Solution:
𝑥𝑛 1(40) 40
= = = 10
4 4 4
<cf = 4
f =6
i =8
lb = 31.5
n=40
2. P36 Solution:
𝑥𝑛 36(40) 1440
= 100 = 100 =14.4
100
<cf = 10
f = 12
i =8
lb = 39.5
=39.5 + (35.2
12
)
= 39.5 + 2.9
= 42.4
Class Interval f <cf
16-23 1 1
24-31 3 4
32-39 6 10
40-47 12 22
48-55 D8 10 32
56-63 8 40
n=40
3. D8 Solution:
𝑥𝑛
10
= 8(40)
10
= 320
10
= 32
<cf = 22
f = 10
i =8
lb = 47.5
=47.5 + 8
= 55.5
computation of the quartile deviation, a measure of variability. The quartiles are also used to determine
who of a group belong to the lower quartile, middle 50%, or upper quartile.
USES OF PERCENTILE
The percentile is computed and used when:
1. A scaled group of scores are to be divided into 100 equal parts of subgroups.
2. percentile bands are needed, that is, when a group of scores is to be divided into a number of subgroups
with equal or unequal number of scores in the subgroups. This is used especially in the transmutation of
raw scores into school marks or grades.
3. percentile ranks of scores are desired. When there is a large number of scores, percentile ranks are
more useful than ordinal ranks.
Measures of Variability or Dispersion are measures of the average distance of each observation from the
center of the distribution. They measure the homogeneity or heterogeneity of a particular group.
MALES
! ! ! ! ! !
60 70 80 90 95 100
FEMALES
! ! ! ! !
60 70 80 90 100
GE 4- Mathematics in the Modern World La Carlota City College
Notice that the grades of the males are far apart from each other, while the grades of the female are more
compressed or clustered together. Thus the measure of the center of the distribution is of little help in describing
and comparing these two sets of data. By getting the average distance of each item from the center of the
distribution, the group can be described more completely and, likewise, similarities and differences can be easily
identified.
RANGE
The range refers to the difference between the highest and the lowest score. If the highest score is 25
and the lowest score in a distribution is 10, then the range is equal to 25-10 = 15. This range is specifically
called as the exclusive range. If the difference between the exact lower limit of the lowest score and the exact
upper limit of the highest score is solved, the result is called the inclusive range. Using the same example, we
get 16 from 25.5-9.5. The inclusive range is simply determined by adding 1 to the exclusive range. Generally,
the exclusive range is used for ungrouped data while the inclusive range is used for grouped data.
The range is the easiest and simplest to determine among the measures of variability because it depends
only on the pair of extreme values. However, it is also the most unstable because its value easily fluctuates with
the change in either of the highest or lowest scores. It is also considered the most unreliable because it does not
give the dispersion or spread of the scores in between the two extreme values. A more reliable measure should
involve all the values in a distribution to provide us an adequate spread of all the scores from the average.
Example 1. Find the range of the grades in Math of the two groups of students in the preceding example.
Male: 100-60 = 40
Female: 83-79=4
The range of grades of the male group is 40 while that of female group is 4. This shows that the grades of the
males are scattered while the grades of the female group are close to each other. It shows further that females
are more homogeneous than the males in their math ability.
Mean Absolute Deviation is the average of the summation of the absolute deviation of each observation
from the mean.
x is a value or score from the raw data
Σ/x−𝑥̅ /
MAD= 𝑁 , where 𝑥̅ is the mean
N is the total number of scores
Example 1. A. Find the mean absolute deviation of the male group in example 1.
Solution: The mean of the male group is 81.
Score Mean /x-𝑥̅ /
x 𝑥̅
GE 4- Mathematics in the Modern World La Carlota City College
70 81 11
95 81 14
60 81 21
80 81 1
100 81 19
Σx=405 Σ/x-𝑥̅ / =66
Σ/x−𝑥̅ / 66
MAD= = = 13.2
𝑁 5
Σ/x−𝑥̅ / 6
MAD= 𝑁 = 5 = 1.2
The male group has a MAD of 13.2 while the female group has 1.2. These results confirm our earlier findings
using the range that the female group is more homogeneous than the male group.
VARIANCE
Variance is the average of the squared deviation from the mean. Formulas for finding the variance for
ungrouped data are shown below:
Σ/x−𝑥̅ /2 Σ/x−𝑥̅ /2
Population Variance : Vp= Sample variance Vs=
𝑁 𝑛−1
STANDARD DEVIATION
Standard Deviation is the square root of the average deviation from the mean, or simply the square root
of the variance.
Σ/x−𝑥̅ /2 Σ/x−𝑥̅ /2
Population Standard Deviation: SDp= √ Sample Standard Deviation: SDs = √
N n−1
GE 4- Mathematics in the Modern World La Carlota City College
Example 4. Find the variance and the standard deviation of male groups in Example 1.
Male group:
x 𝑥̅ x-𝑥̅ (x-𝑥̅ )2
70 81 -11 121
95 81 14 196
60 81 -21 441
80 81 -1 1
100 81 19 361
Σx=405 Σ (x-𝑥̅ )2= 1120
a. Treating the data as population, the variance and the standard deviation is
1120 1120
Vp= 5 = 224 square units Sample variance Vs= 4 = 280 square units
• Find the variance and the standard deviation of the female group.
An equivalent formula gives another way of computing the standard deviation. It does away with
computing for the mean and the deviation of the scores. We can use the equivalent formula below.
2 (Σx)2
Σx −
SDs = √ n − 1n
Male Group:
x X2
70 4900
95 9025
60 3600
80 6400
100 10,000
Σx=405 Σx2= 33925
2 (405)2
2 − (Σx) 164025
SDs = √Σx n
= √33925− 5
= √
33925−
5
= √
33925 − 32805
= √
1120
= √280 =
n−1 5−1 4 4 4
16.73 units
variance = 280, standard deviation = 16.73
GE 4- Mathematics in the Modern World La Carlota City College
Formula:
Σf(x−𝑥̅ ) 2
Vp = Σf (x-𝑥̅ )2 / N and SDp =√ N
̅̅̅ 2
Σf(x−𝑥)
Vs= Σf (x-x)2 / n-1 and SDs = √ n−1
Example: The grouped data which gives the wages per day of the laborers in a certain construction.
Activity 3:
1. A group of second-year Business Administration students took the qualifying examination for admission to the
course Bachelor of Science in Accounting. The results are as follows:
64 50 59 73 88 75 78 52 75 53
54 95 68 58 71 57 76 37 49 90
50 49 90 87 42 84 68 56 70 60
44 61 71 66 74 71 39 34 68 65
60 75 74 85 65 46 64 77 77 78
a. Prepare a frequency distribution using 7 classes starting with 34.
b. Include the columns of % relative frequency, less than cumulative frequency, and greater than cumulative
frequency.
2. The data below is the frequency distribution of the Intelligence Quotient (IQ) of 200 students taking up
Behavioral Psychology and Business Statistics at a certain university.
Classes frequency (f)
131-135 7
126-130 10
121-125 14
116-120 35
111-115 42
106-110 21
101-105 18
96-100 29
91-95 10
86-90 14
N=200
Find the following:
1. Class interval size
2. Lower class boundary of the highest class interval.
3. Number of students whose IQ is greater than 115.5
4. Number of students whose IQ is less than 96
5. Lowest upper class limit
6. Upper class boundary of the 3rd class interval
7. Upper class limit of the 5th class interval
8. Number of students whose IQ is between 126-130
9. Highest upper class limit
10. Class mark of the 4th class of the highest frequency
3. The following are the test scores obtained by III-1 students in Statistics. Compute the mean using the:
a. Classmark formula
b. Coded formula
c. What is the average score obtained by the students?
Class Interval f x fx <cf
20-24 4
25-29 6
30-34 7
35-39 10
40-44 5
45-49 8
N=
GE 4- Mathematics in the Modern World La Carlota City College
4. The following data give the time (in minutes) taken to commute from home to school for 20 students of LCCC.
10 50 64 33 48 5 11 23 37 26
26 32 17 7 13 19 29 43 21 22
Find the mean, median, mode, Q1, D5, P70, SDp, Vp, SK, KU
5. Below is the frequency distribution showing the result of an IQ test of a group of students in a certain college.
Classes frequency(f)
80-85 2
86-91 8
92-97 19
98-103 21
104-109 25
110-115 52
116-121 12
122-127 11
6. The NSAT scores of 12 students in a certain college were taken and are shown below.
86, 95, 84, 87, 91, 90, 99, 84, 83, 88, 96, 92
Determine:
a. Mean b. Median c. Q3 d. P40
7. The IQ’s of 5 members of the family are 108, 112, 127, 118 and 113.
Find the
a. Range b. Variance c. Standard Deviation
8. Below is a frequency distribution showing the result of an IQ test of a group of students in a certain college.
Classes f
80-85 2
86-91 8
92-97 19
98-103 21
104-109 25
110-115 52
116-121 12
122-127 11
Find:
a. Range b. Mean Absolute Deviation
c. Population Variance d. Population Standard Deviation