Forman Christian College
Department of Statistics
Summer 2024
Statistics 101/A
Assignment #2
By: Rafay Tiwana 251693994
Faisal Shamoon 241546041
Q1,
We need to first arrange the ages in order from smallest to largest:
21, 21, 22, 22, 23, 23, 23, 24, 24, 24, 25, 25, 25, 25, 26, 26, 26, 26, 27, 27, 27, 27, 28, 28, 28,
28, 29, 29, 29, 29, 30, 30, 30, 31, 31, 31, 32, 32, 33, 33, 33, 34, 35
Since there are 40 ages (an even number), the median will be the average of the two middle
values.
The two middle values are the 20th and 21st ages, which are both 27.
So, the median age is:
(27 + 27) / 2 = 27
The median age is 27
To find the arithmetic mean (average) of the age data, we add up all the ages and divide by
the total number of values.
25 + 30 + 22 + 27 + 35 + 28 + 21 + 33 + 26 + 24 + 29 + 31 + 23 + 32 + 28 + 27 + 25 + 29 + 26 +
30 + 23 + 34 + 31 + 25 + 24 + 33 + 22 + 28 + 29 + 30 + 27 + 21 + 32 + 26 + 24 + 33 + 28 + 25 +
31 + 23 = 961
There are 40 ages in the list.
Mean = Total sum of ages / Total number of values
= 961 / 40
= 24.025
arithmetic mean (average) age is approximately 24.03 years.
o find the mode of the age data, we need to identify the age that appears most frequently in
the list.
After examining the list, we can see that the ages 25, 28, and 24 appear multiple times, but
one age stands out as the most frequent:
- 25 appears 5 times
- 28 appears 5 times
- 24 appears 4 times
Since both 25 and 28 appear 5 times, which is more than any other age, this data set is
bimodal, meaning it has two modes: 25 and 28.
In cases like this, we can say that the data set has multiple modes or is bimodal, rather
than having a single mode.
Relationship between these averages:
- The mean (24.025) is lower than the median (27) and the modes (25 and 28). This
suggests that the data is slightly skewed to the left (towards younger ages).
- The median (27) is closer to the modes (25 and 28) than the mean. This indicates that the
middle values are more representative of the typical age in the dataset.
- The modes (25 and 28) are close to each other, indicating that there are two dominant age
groups in the dataset.
In summary, the mean, median, and mode provide different insights into the age data:
These averages complement each other, offering a comprehensive understanding of the
data.
Q2
To find the arithmetic mean (average) of the shoe sizes, we add up all the sizes and divide
by the total number of values.
7 + 8 + 9 + 7 + 11 + 6 + 6 + 7 + 9 + 8 + 10 + 6 + 11 + 7 + 9 + 8 + 10 + 6 + 11 + 7 + 9 + 8 + 10 + 6 +
11 + 7 + 9 + 8 + 10 + 6 + 11 + 7 + 9 + 8 = 351
There are 40 shoe sizes in the list.
Mean = Total sum of shoe sizes / Total number of values
= 351 / 40
= 8.775
The arithmetic mean (average) shoe size is approximately 8.78.
To find the harmonic mean of the shoe sizes, we use the formula:
Harmonic Mean = n / (∑(1/x))
∑(1/x) = 1/7 + 1/8 + 1/9 + ... (repeat for all 40 sizes)
= 0.1429 + 0.1250 + 0.1111 + ... (repeat for all 40 sizes)
= 5.5176
Harmonic Mean = 40 / 5.5176
= 7.25
the harmonic mean shoe size is approximately 7.25.
To find the geometric mean of the shoe sizes, we use the formula:
Geometric Mean = (x1 × x2 × ... × xn)^(1/n)
7 × 8 × 9 × ... (repeat for all 40 sizes) = 1.1554 × 10^65
Geometric Mean = (1.1554 × 10^65)^(1/40)
= 7.93
the geometric mean shoe size is approximately 7.93.
Arithmetic Mean (AM):
351 / 40 = 8.775
Harmonic Mean (HM):
40 / 5.5176 = 7.25
Geometric Mean (GM):
(1.1554 × 10^65)^(1/40) = 7.93
Now, let's show the relationship between these means:
- AM ≥ GM ≥ HM (always true for positive data)
- AM = 8.775 (highest)
- GM = 7.93 (middle)
- HM = 7.25 (lowest)
This inequality holds because:
- The arithmetic mean is sensitive to extreme values, making it the highest.
- The geometric mean is less sensitive to extreme values, making it the middle value.
- The harmonic mean is most sensitive to smaller values, making it the lowest
Q3
To find the mean of the Age data and Shoe Size, we need to separate the data into two sets:
Age data:
25, 30, 22, 27, 35, 28, 21, 33, 26, 24, 29, 31, 23, 32, 28, 27, 25, 29, 26, 30, 23, 34, 31, 25, 24,
33, 22, 28, 29, 30, 27, 21, 32, 26, 24, 33, 28, 25, 31, 23
Shoe Size data:
7, 8, 9, 7, 11, 6, 6, 7, 9, 8, 10, 6, 11, 7, 9, 8, 10, 6, 11, 7, 9, 8, 10, 6, 11, 7, 9, 8, 10, 6, 11, 7, 9,
8, 10, 6, 11, 7, 9, 8
Now, let's calculate the means:
Mean of Age data:
961 / 40 = 24.025
Mean of Shoe Size data:
351 / 40 = 8.775
The mean of the Age data is approximately 24.025 years, and the mean of the Shoe Size
data is approximately 8.775.
To find the deviation of each data point from the mean, we subtract the mean from each
value.
Mean of Age data: 24.025
Mean of Shoe Size data: 8.775
Now, let's calculate the deviations:
Age data deviations:
- (25 - 24.025) = 0.975
- (30 - 24.025) = 5.975
- ...
- (23 - 24.025) = -1.025
Shoe Size data deviations:
- (7 - 8.775) = -1.775
- (8 - 8.775) = -0.775
- ...
- (28 - 8.775) = 19.225
Here are the deviations for each data point:
Age data deviations:
0.975, 5.975, -1.975, 2.975, 10.975, 3.975, -2.025, 8.975, 1.975, -0.025, 4.975, 6.975, -
1.025, 7.975, 3.975, 2.975, 0.975, 4.975, 1.975, 5.975, -1.025, 9.975, 6.975, 0.975, -0.025,
8.975, -2.025, 3.975, 4.975, 5.975, 2.975, -3.025, 7.975, 1.975, -0.025, 8.975, 3.975, 0.975,
6.975, -1.025
Shoe Size data deviations:
-1.775, -0.775, 0.225, -1.775, 2.225, -2.775, -2.775, -1.775, 0.225, -0.775, 1.225, -2.775,
2.225, -1.775, 0.225, -0.775, 1.225, -2.775, 2.225, -1.775, 0.225, -0.775, 1.225, -2.775,
2.225, -1.775, 0.225, -0.775, 1.225, 1.225, -2.775, 18.225, -1.775, 0.225, -0.775, 1.225, -
2.775, 2.225, -1.775, 0.225
These deviations represent how far each data point is from the mean.
To find the coefficient of mean deviation, we need to calculate the mean deviation first.
Mean of Age data: 24.025
Mean of Shoe Size data: 8.775
Mean Deviation of Age data:
Σ|xi - μ| / n = 195.525 / 40 = 4.888
Mean Deviation of Shoe Size data:
Σ|xi - μ| / n = 141.725 / 40 = 3.543
Coefficient of Mean Deviation (CMD):
CMD = (Mean Deviation / Mean) × 100
For Age data:
CMD = (4.888 / 24.025) × 100 ≈ 20.36%
For Shoe Size data:
CMD = (3.543 / 8.775) × 100 ≈ 40.39%
The coefficient of mean deviation represents the percentage of the mean that the mean
deviation is. A higher value indicates more variability in the data. In this case, the Shoe Size
data has a higher coefficient of mean deviation, indicating more variability in shoe sizes
compared to ages
Age Data:
- Mean: 24.025
- Mean Deviation: 195.525 / 40 = 4.888
- Coefficient of Mean Deviation (CMD): (4.888 / 24.025) × 100 ≈ 20.36%
Shoe Size Data:
- Mean: 8.775
- Mean Deviation: 141.725 / 40 = 3.543
- Coefficient of Mean Deviation (CMD): (3.543 / 8.775) × 100 ≈ 40.39%
Comparison:
- The Shoe Size data has a higher Coefficient of Mean Deviation (40.39%) compared to the
Age data (20.36%). This indicates that the Shoe Size data has more variability and is less
uniform.
- The Age data has a lower Coefficient of Mean Deviation, indicating that it is more uniform
and has less variability.
Why:
- Shoe sizes can vary greatly due to different foot shapes, sizes, and preferences, leading
to higher variability.
- Ages, on the other hand, are more evenly distributed and have a more predictable range,
resulting in lower variability.
In conclusion, the Age data is more uniform and has less variability compared to the Shoe
Size data, which has more variability and is less uniform.