Continuous Probability Distributions Explained
Continuous Probability Distributions Explained
x 74 x 72
b) No. The equation has only one solution.
4 6
a) 5! = 5(4)(3)(2)(1) 3! 10! 7!
= 120 b) P c) P2 d) C4
(3 3)! (10 2)! 4!(7 4)!
3 3 10 7
3! 10 9 8! 7 6 5 4!
0! 8! 4!3!
3(2)(1) 10 9 765
1 90 3 2 1
6 35
Order does not matter. Frank can choose his 5 rides in 17C5, or 6188 ways.
a) Order matters. Wayne can arrange the 12 books in 12P12, or 479 001 600 ways.
b) Wayne can arrange the books in 3! × 3! × 3! × 3!, or 1296 ways, if they are clumped in
groups of 3 for each grade in increasing order from left to right.
There are a total of 15 gumballs with 5 green gumballs. The probability of Sam choosing green is
5 1
, or .
15 3
Order matters. The six student can be lined up in 6P6, or 720 ways. Only 1 of those ways is in
alphabetical order. So, the probability that they are lined up in alphabetical order by first name is
1
.
720
These are independent trials. For a number greater than 9, you could roll a 10, 11, or 12 with the
3 1
12-sided die. So, the probability of rolling a number greater than 9 is , or . This is a
12 4
1 3
binomial distribution with p = and q = . The probability that a roll comes up greater than 9
4 4
4 3
1 3
exactly four times in seven rolls is 7 C4 , or about 0.0577.
4 4
12 1
These are independent trials. The probability of the barrier being down is , or . This is a
60 5
1 4
binomial distribution with p = and q = . The probability that exactly 20 cars will find the
5 5
20 80
1 4
barriers down in 100 cars is 100 C20 , or about 0.0993.
5 5
There are 26 students and 4 have red hair. Four are to be selected: n(S) = 26C4.
For the probability of 4 red heads, n(A) = 4C4 × 22C0.
C C
P ( A) 4 4 22 0
26 C4
1
14 950
6.689 105
The probability that all four have red hair is about .0067%.
Of the 1000 students at Eastdale Secondary School, 420 are boys and 580 are girls.
Ten are to be selected: n(S) = 1000C10.
For the probability of 5 boys and 5 girls, n(A) = 420C5 × 580C5.
1000 C10
0.2170
The probability that an equal number of boys and girls will receive a free sample is about 21.7%.
1
a) These are independent trials. The probability of a 50%-discount card is . This is a
100
1 99
binomial distribution with p = and q = . The probability that exactly 3 customers
100 100
3 247
1 99
received a 50%-discount card in 250 customers is 250 C3 , or about 0.2149.
100 100
b) The probability that a random variable assumes a value between a and b is the area under the
curve between a and b. The probability that it takes Chris between 9 min and 11.5 min to change
a tire equals the shaded area under the graph from 9.0 min to 11.5 min.
P(9 x 11.5) area under graph
0.25(11.5 9.0)
0.625
The probability that it takes Chris between 9 min and 11.5 min to change a tire is 0.625.
a) and b)
c) The mean cubit length in the class appears to be around 523 mm.
d) All outcomes are not equally likely. So, this is not a uniform distribution.
The number of guests that occupy the hotel each day is discrete data, while the time a guest waits
for an elevator is continuous data. It is possible to list all values of the discrete distribution, since
these would be values from 0 up to the maximum capacity of the hotel. It is not possible to list all
values of the continuous distribution, since time can be recorded in fractions of a second.
Answers may vary. The various height measurements could be caused by several things. For
example, Maya standing differently each time, bad measuring technique, and misreading the
measuring device. This is most likely measurement error, not bias.
Answers may vary. The probability that a variable falls within a range of values is equal to the
area under the probability density graph for that range of values. The area method cannot be used
for single values of a continuous variable, only for a range of values. A continuous random
variable can take on an infinite number of values. The probability that it will equal a specific
value is always zero.
a) Mass represents a continuous variable. This will not result in a discrete distribution.
c) Barometric pressure represents a continuous variable. This will not result in a discrete
distribution.
a) The number of students with blue eyes represents a discrete variable and would not result in a
continuous distribution.
b) Weight represents a continuous variable and would be expected to result in a continuous
distribution.
c) The number of cartons of milk represents a discrete variable and would not result in a
continuous distribution.
d) The number of defective tablets represents a discrete variable and would not result in a
continuous distribution.
a) The distribution does not appear to be uniform because each waist size interval is not equally
likely.
c) If a data value falls on the boundary between two intervals, it is usually placed in the lower
interval. So, the additional waist size of 38 should go in the 36–38 interval.
a) The data are difficult to analyse in this form. It is not obvious whether the distribution is
uniform or not.
b) Use a table to determine the frequency for each interval. If all frequencies are equal, then the
distribution is uniform.
Volume of Sample (mL) Frequency
54−55 4
55−56 4
56−-57 4
57−58 4
58−59 4
59−60 4
60−61 4
61−62 4
62−63 4
63−64 4
64−65 4
This distribution is uniform.
a) Use a table to determine the frequency for each interval. If all frequencies are equal, then the
distribution is uniform.
Pressure (psi) Frequency
2995−2997 8
2997−2999 6
2999−3001 7
3001−3003 7
3003−3005 4
The distribution is not uniform.
c) Answers may vary. A possible reason for why there are no values with decimal places is that
the gauge used to measure the pressure only has a whole-number scale.
a) The probability of 30 min or less equals the shaded area under the graph from 25.0 min to
30.0 min.
P(25.0 x 30.0) area under graph
0.05(30.0 25.0)
0.25
The probability a contestant will finish the triathlon in 30 min or less is 0.25.
b) The probability of 30 min to 40 min equals the shaded area under the graph from 30.0 min to
40.0 min.
P(30.0 x 40.0) area under graph
0.05(40.0 30.0)
0.5
The probability a contestant will finish the triathlon in 30 min to 40 min is 0.5.
c) This type of distribution was used because all times are equally likely.
a) From the diagram, there are 88 keys including both black and white keys.
b) Answers may vary. I do not expect the distribution to be uniform, because I doubt that all
keys will be used in a piece of music.
c) Answers may vary. Since each key represents a specific note (frequency), the distribution in
part b) is discrete.
b) If a trombone player plays a musical composition, he plays specific notes. I would expect the
distribution to be discrete.
a) There are six possible lengths for the sardines. A quick inspection shows that there are six
occurrences of 99 mm. So, it is not possible for the 24 sardine sample to be a uniform
distribution.
c)
d) The probability of a length less than or equal to 98 mm equals the shaded area under the
graph from 95.0 mm to 98.0 mm.
P(95.0 x 98.0) area under graph
0.2(98.0 95.0)
0.6
The probability that a sardine has a length less than or equal to 98 mm is 0.6.
The area method cannot be used for single values of a continuous variable. So, Jon is incorrect.
The probability that a variable falls within a range of values is equal to the area under the
probability density graph for that range of values. The area shaded is from 2.5 to 3.5. So, Sunita is
correct.
a) The low and high cutoff values for this distribution is 5%, or 95 kΩ to 105 kΩ.
d) The low and high cutoff values for 0.25%, or 99.75 kΩ to 100.25 kΩ.
P(99.75 x 100.25) area under graph
0.1(100.25 99.75)
0.05
The probability that a given resistor is within .25% of the stated resistance value is 0.05.
b) The data ranges from 64.6 to 80.2. So, I chose intervals of width two, starting at 64.5.
c)
Speed (km/h) Frequency
64.5–66.5 4
66.5–68.5 4
68.5–70.5 3
70.5–72.5 5
72.5–74.5 5
74.5–76.5 3
76.5–78.5 5
78.5–80.5 1
d)
1
g) Answers may vary. Since only one speed was recorded over 80 km/h, my estimate is , or
30
about 0.033.
Answers may vary. I think that hair colour should be considered continuous, since there are
unlimited numbers of shades between blonde, red, brown, and black.
b) Three instruments that can play any frequency over a range are bass, cello, violin.
a) To determine the probability, read the relative frequency from the table for waist size 30–32
in. The probability that a customer has a waist size between 30 and 32 in. is 0.175.
b) To determine the probability, add the relative frequencies from 36 in. to 42 in.
0.16 + 0.045 + 0.000 = 0.205
The probability that a customer has a waist size of more than 36 in. is 0.205.
c) To determine the probability, add the relative frequencies from 30 in. to 36 in.
0.175 + 0.295 + 0.300 = 0.770
The probability that a customer has a waist size between 30 in. and 36 in. is 0.770.
d) Probabilities for a continuous distribution cover a range of values. The probability that a
waist size will be exactly 38 in. would require an interval width of 0 in., which is not available on
the table. The probability that a waist size will be exactly 38 in. is zero.
b) Choose –99999 as the lower bound and an upper bound of 30. Use the values for the mean
and standard deviation that were calculated in part a): normalcdf(–99999, 30, 35.2, 9.567).
The probability that Kunal kicks a distance of less than 30 yd is about 0.2934.
Answers may vary. As Kunal practises and gains more skill, I expect the mean and standard
deviation to change. As he improves and becomes more consistent, I expect the mean to increase
and the standard deviation to decrease.
The z-scores for a normal distribution follow a normal distribution themselves, with a mean of 0
and a standard deviation of 1. From the graph, P(z < –5) is located far in the left tail. So, the
probability that far from the central peak is essentially zero.
Use the formula for sample z-score with x = 13.8, x = 10.2, and s = 2.4.
xx
z
s
13.8 10.2
2.4
1.5
The z-score for a trip that takes 13.8 min is 1.5. Answer C.
Use the formula for sample z-score with z = –0.5, x = 10.2, and s = 2.4.
xx
z
s
x 10.2
0.5
2.4
0.5(2.4) x 10.2
1.2 x 10.2
x9
The trip with a z-score of –0.5 took 9 min. Answer A.
A trip that is twice as long as the mean is 2(10.2), or 20.4 min. Recall that the probability that a
variable takes on a specific value is [Link] probability of a trip taking exactly 20.4 min is 0.
Answers may vary. Data collected one week could be different from another because of the path
chosen, traffic lights, or weather condition. Roberta could obtain more reliable values for the
mean and standard deviation by combining the data from the two weeks.
a)
b) The mean life of these light bulbs appears to be around 437.5 days.
c)
Lifetime (days) Frequency Relative Frequency
300–325 2 0.004
325–350 15 0.030
350–375 38 0.076
375–400 55 0.110
400–425 91 0.182
425–450 94 0.188
450–475 73 0.146
475–500 68 0.136
500–525 40 0.080
525–550 14 0.028
550–575 9 0.018
575–600 1 0.002
d) To determine the probability, add the relative frequencies from 300 days to 400 days.
0.004 + 0.030 + 0.076 + 0.110 = 0.220
e) Answers may vary. Replace the light bulbs every 437.5 days, the estimated mean, to be
reasonably sure that there would never be a burned out light bulb.
a)
Speed of Ball
(km/h) Frequency
49–54 3
54–59 10
59–64 19
64–69 7
69–74 1
b)
c)
d) Answers may vary. While the distribution is centred around a central value, it does not drop
off symmetrically to the left and right.
g) To determine the probability, add the relative frequencies from 59 km/h to 69 km/h.
0.475 + 0.175 = 0.65
The probability that a given ball will launch at a speed of 59 km/h to 69 km/h is 0.65.
a)
Speed of Ball
(km/h) Frequency
49–51 2
51–53 1
53–55 1
55–57 4
57–59 5
59–61 8
61–63 10
63–65 3
65–67 2
67–69 3
69–71 1
b)
e) Answers may vary. I think the smaller interval width made it easier to estimate the mean.
f) Answers may vary. The smaller interval width gives a clearer picture of the actual
distribution because it better approximates the shape of the frequency distribution.
a)
Horizontal Error
(cm) Frequency
(–10)–(–8) 1
(–8)–(–6) 3
(–6)–(–4) 6
(–4)–(–2) 6
(–2)–0 8
0–2 3
2–4 4
4–6 4
6–8 1
d) Answers may vary. A misadjusted sight with a bias to the left would result in more negative
values.
b) See graphing calculator screen in part a). The standard deviation, Sx, is about 2.3612 cm.
c) Answers may vary. If a variable is expected to follow a normal distribution, you can take a
representative sample. However, there is not enough data to predict whether the distribution is
normal. The data given shows that distribution may be centred, but it is unclear whether it will
drop off symmetrically to the left and right.
Tree Height (cm) Frequency
30–32 3
32–34 2
34–36 7
36–38 7
38–40 1
a) Calculate the z-score for each student. Determine which student is performing at a higher
number of standard deviations from the respective means.
b) Answers may vary. The university must assume that the tests are comparable in material and
level of difficulty.
c)
Region A: Use x = 400, x = 350, and s = 35. Region B: Use x = 67, x = 62, and s = 5.
xx xx
z z
s s
400 350 67 62
35 5
1.4286... 1
Since 1.429 > 1, the student from region A performed at a level that is further above average than
the student from region B.
b) See graphing calculator screen in part a). The standard deviation, Sx, is about 0.0196.
d) Use the table to determine the associated probability for a z-score of –1.28 as 0.1003 and the
associated probability for a z-score of 0.77 as 0.7794.
P(8.38 ≤ X ≤ 8.42) = P(X ≤ 8.42) − P(X ≤ 8.38)
= 0.7794 – 0.1003
= 0.6791
Then, the probability that a given manufactured piston falls outside the acceptable range is 1 –
0.6791, or 0.3209.
Use the table to determine the associated probability for a z-score of –0.4 as 0.3446 and the
associated probability for a z-score of 0.4 as 0.6554.
P(4680 ≤ X ≤ 5720) = P(X ≤ 5720) − P(X ≤ 4680)
= 0.6554 – 0.3446
= 0.3108
Then, the probability that the release will be within 10% of the mean is 0.3108.
c) Marks this year might be lower than expected because the calibre of the students has
declined.
a) Consider the formula for the mean of grouped data, which is the same as the sum of the
product of the midpoint of each interval, mi, and the relative frequency of each interval, rfi.
x mi rfi
491(0) 493(0) 495(0.10) 509(0)
501.08
The mean soft drink volume is 501.08 mL.
c) The value of the mean makes sense. It occurs in the centre of the bell-shaped frequency
polygon.
a) Three standard deviations below the mean results in a z-score of −3, while three above the
mean results in a z-score of +3. Use the table to determine the probability that a value lies in the
range. Note: The table only contains z-score values from –2.99 to 2.99.
P(3 z 3) P( z 2.99) P( z 2.99)
0.9986 0.0014
0.9972
The probability that a value lies within three standard deviations of the mean is about 99.7%.
Answers may vary. Determine the z-score for a length of 4.7 cm.
Use x = 4.7, x = 5, and s = 0.1.
xx
z
s
4.7 5.0
0.1
3
The probability that the length of a bolt is 4.7 cm or less is only about 0.1%, so this would be a
surprising value.
Since data collected from a large sample of people or naturally occurring phenomena usually
have a normal distribution, it is reasonable to expect the heights or students in a grade 3 class, the
mass of peanut butter in a sample of jars, and the distance that a person can throw a football
follow a normal distribution. Answer D.
Answers may vary. A quality control engineer would be interested in the mean and standard
deviation of the washers to ensure that they will actually fit onto the standard bolts. In particular,
the internal diameter of the washer must not be smaller the exterior diameter of the bolt shaft.
b)
The mean, x , is about 701 mm and the standard deviation, Sx, is about 20.1 mm.
c) The random sample mean and standard deviation are fairly close to the underlying normal
distribution.
g) The larger the sample, the closer the sample measures are to the underlying normal
distribution.
b) The distance needed to span one and one-half octaves is 24.6 cm.
Determine the z-score for a hand span of 24.6 cm.
Use x = 24.6, x = 21.8, and s = 2.4.
xx
z
s
24.6 21.8
2.4
1.1666...
Use the table to determine the associated probability for a z-score of 1.17 as 0.8790.
Determine the probability that a value lies above this range.
1 – P(X < 24.6) = 1 – 0.8790
= 0.121
The probability that a student could play one and one-half octaves is 0.121.
b) Use the table to determine the associated probability for a z-score of –2 as 0.0228. Then, the
expected number of watches that the company will replace is 100 000(0.0228), or 2280. So, the
company should budget $5(2280), or $11 400 to replace watches under warranty.
a) A score of 60% is one standard deviation below the mean and a score of 80% is one standard
deviation above the mean. For a normal distribution, 68% of the data values fall within one
standard deviation of the mean. So, 200(0.68), or 136 students are expected to score between 60%
and 80%.
b) A score of 50% is two standard deviations below the mean and a score of 90% is two
standard deviations above the mean. For a normal distribution, 95% of the data values fall within
two standard deviations of the mean. So, 200(0.95), or 190 students are expected to score
between 50% and 90%.
c) Use the table to determine the associated probability for a z-score of –2 as 0.0228. So,
200(.0228), or about 5 students would be expected to score below 50%.
For a normal distribution, 95% of the data values fall within two standard deviations of the mean.
0.0001
So, the standard deviation needed to meet this requirement is , or 0.000 05 in.
2
a) Answers may vary. If the data set is reasonably large and the data fall into a symmetric bell
shape, then it seems reasonable to use a normal distribution to model the discrete data.
b) Answers may vary. I think a reasonable minimum number of hot dogs and buns is the mean,
120. This would cover 50% of the data in the analysis.
The mean, x , is about 503 g and the standard deviation, Sx, is about 2 g.
c) The expected percent of the data within one standard deviation is 68%. So, these values are
more clustered around the mean.
a) Determine the standard deviation. From the table, 0.05 has a z-score of –1.65.
Use z = –1.65, x = 200, and x = 220.
xx
z
s
200 220
1.65
s
1.65s 20
s 12.1212...
The standard deviation of the amount of corned beef is about 12.1 g.
b) To ensure that no more than 0.5% of the sandwiches contain less than 200 g of corned beef,
Rudy could buy a better slicing machine.
Determine the new standard deviation. From the table, 0.005 has a z-score of –2.58.
Use z = –2.58, x = 200, and x = 220.
xx
z
s
200 220
2.58
s
2.58s 20
s 7.7519...
The standard deviation of the amount of corned beef is now about 7.8 g.
To ensure that no more than 0.5% of the sandwiches contain less than 200 g of corned beef, Rudy
could increase the slicing machine mean.
Determine the mean. From the table, 0.005 has a z-score of –2.58.
Use z = –2.58, x = 200, and s = 12.1.
xx
z
s
200 x
2.58
12.1
31.218 200 x
x 231.218
The mean amount of corned beef is now about 231.2 g.
c) Answers may vary. Most likely increasing the slicing machine mean will be more cost
effective than buying a new machine.
a) For two standard deviations from the mean, k = 2. According to Chebyshev’s Theorem, no
1
more than 2 , or 25% of the values lie more than two standard deviations from the mean.
2
b) Since 95% of the data values lie within two standard deviations of a normal distribution, then
5% of the data lie outside this range. Chebyshev’s Theorem does not exactly agree, but could
since the 25% value is a maximum.
c) The proportion of values that must lie within k standard deviations of the mean is given by
1
1 2 .
k
a) A confidence level of 99% has a z-score of 2.576. Use the formula for the margin of error
with z = 2.576, p = 0.75, and n = 100.
p (1 p)
Ez
n
0.75(1 0.75)
2.576
100
0.112
The margin of error at a 99% confidence interval is about 11.2%.
c) Answers may vary. Hockey Night in Canada is watched by 75% of households. This estimate
is considered correct within 11.2%, 99 times out of 100.
a) A confidence level of 90% has a z-score of 1.645. Use the formula for the margin of error
with z = 1.645, p = 0.13, and n = 400.
p(1 p)
Ez
n
0.13(1 0.13)
1.645
400
0.028
The margin of error at a 90% confidence interval is about 2.8%.
b) Use the rearranged form of the margin of error formula with z = 1.645, p = 0.13, and
E = 0.014.
998 1234 1523 1760 937 1193 996 1002 986 1285 1163 1716
a) x
12
1232.75
x
n
420
12
121.24
The mean of the sample is 1232.75 h, and the standard deviation of the sample is about 121.24 h.
No. The confidence level is the probability that a particular statistic is within the range indicated
by the margin of error. The confidence level’s related z-score is used to calculate the margin of
error.
The lower end of the range is 5.6% – 1.4%, or 4.2%. The upper end of the range is 5.6% + 1.4%,
or 7%. The 90% confidence interval for the mean defective tablets within one year is 4.2% to 7%.
This is the range of possible percents of defective tablets.
Answers may vary. This would help the manufacturer budget for returns and provide information
on the reliability of the manufacturing process.
A 90% confidence interval can be stated as 9 times out of 10. A 99% confidence interval can be
stated as 99 times out of 100.
a) For a 90% confidence interval, z = 1.645. The values will be about 1.6 standard deviations
from the mean.
b) For a 95% confidence interval, z = 1.960. The values will be about 2 standard deviations from
the mean.
c) For a 99% confidence interval, z = 2.576. The values will be about 2.6 standard deviations
from the mean.
17
The confidence level for this poll is , or 85%. Answer B.
20
σ
σx
n
5000
100
500
The expected standard deviation of the sample means is 500 km.
a) A confidence level of 99% has a z-score of 2.576. Use the formula for the margin of error
with z = 2.576, p = 0.82, and n = 25.
p (1 p)
Ez
n
0.82(1 0.82)
2.576
25
0.198
The margin of error at a 99% confidence interval is about 19.8%.
b) The lower end of the range is 82% – 19.8%, or 62.2%. The upper end of the range is 82% +
19.8%, or 101.8%. The 99% confidence interval for the exam marks is 62.2% to 101.8%.
Answers may vary. For the original poll, the lower end of the range is 34% – 3.4%, or 30.6%.
The upper end of the range is 34% + 3.4%, or 37.4%. The 95% confidence interval for the
average percent of support is 30.6% to 37.4%. The other two polls are both outside of this range,
suggesting that support is not consistent.
c) Answers may vary. No. The company is not justified in claiming the line contains twice the
filling. The Double Crème line contains twice the filling when compared to approximately the
lower half of the Single Crème line distribution. So, for only less than half of the cookies is the
claim valid.
For a confidence level of 99%, z = 2.576. Use the margin of error formula with E = 0.105 and σ =
0.05.
σ
Ez
n
σ2
E2 z2
n
σ2
n z2 2
E
0.052
2.5762
0.1052
1.5047....
The number of businesses surveyed was about 2.
For a confidence level of 95%, z = 1.960. Use the margin of error formula with E = 1.5 and σ =
5.9.
σ
Ez
n
σ2
E2 z2
n
σ2
n z2 2
E
5.92
1.9602
1.52
59.4338....
Approximately 59 patients were in the study.
a) A confidence level of 95% has a z-score of 1.960. Use the formula for the margin of error
with z = 1.960, p = 0.92, and n = 300.
p (1 p)
Ez
n
0.92(1 0.92)
1.960
300
0.031
The lower end of the range is 92% – 3.1%, or 88.9%. The upper end of the range is 92% + 3.1%,
or 95.1%. The 95% confidence interval for the load of tomatoes is 88.9% to 95.1%.
b) The lower end of the range is 0.889(41.992), or about 37.331. The upper end of the range is
0.951(41.992), or about 39.934. The 95% confidence interval for the mass of acceptable tomatoes
is 37.331 tonnes to 39.934 tonnes.
Use the formula for the margin of error with z = 1.960, p = 0.67, and n = 300.
p(1 p)
Ez
n
0.67(1 0.33)
1.960
300
0.053
The lower end of the range is 67%, or about 28.135. The upper end of the range is 67% + 5.3%,
or 72.3%. The interval for the mass of acceptable tomatoes is 28.135 tonnes to 30.360 tonnes.
The lower end of the range is $94.40(28.135), or $2655.91. The upper end of the range is
$94.40(30.360), or $2866. The interval for the worth of acceptable tomatoes is $2655.94 to
$2866.
The caption means that anywhere from 54% to 60% of students would vote for Adam and
anywhere from 48% to 54% of students would vote for Meghan with a 95% confidence level.
b) This results in the standard deviation of the sample mean being smaller than that of the
population.
σ
c) The effect in part b) fits with the formula for the standard deviation of the sample, σ x .
n
The standard deviation of the population is divided by n , thus decreasing the value.
b) μ np σ npq
25(0.25) 25(0.25)(0.75)
6.25 2.165
The mean is 6.25, and the standard deviation is about 2.165.
a) There are 90 socks in the drawer, of which 7 are chosen. The number of trials is less than
10% of the population. The normal approximation is reasonable for this hypergeometric
distribution.
Answers may vary. Any scenario that involves only two outcomes, success (p) and failure (q),
and the number of independent trials is large enough to meet the requirements of np > 5 and
nq > 5. For example, number of heads when flipping a coin or failure rate for quality control.
Answers may vary. Any scenario that involves two outcomes, success (p) and failure (q), in
dependent trials, and the number of dependent trials is less than 10% of the population.
Examples: choosing certain people to be on a committee or the number of a particular type of
card in a five-card hand.
Answers may vary. Using the approximation allows the probabilities of value ranges to be
calculated more easily than with the binomial or hypergeometric formulas.
Answers may vary. In the case of the normal approximation for a binomial distribution, any
scenario where np ≤ 5 or nq ≤ 5. In the case of the normal approximation for a hypergeometric
distribution, any scenario where the number of dependent trials is greater than or equal to 10% of
the population.
13
The probability of drawing a diamond, p, is , or 0.25.
52
μ np σ npq
30(0.25) 30(0.25)(0.75)
7.5 2.372
The mean is 7.5, and the standard deviation is about 2.372. Answer A.
There are 30 red jelly beans out of 200. 15 beans are drawn.
NP n
σ npq
NP 1
30 170 200 15
15
200 200 200 1
1.333
The standard deviation is about 1.333. Answer D.
6
The probability of a double, p, is . To model this situation using a normal distribution, np > 5
36
and nq > 5.
np 5 nq 5
6 24
n 5 n 5
36 36
36 36
n 5 n 5
6 24
n 30 n 3.333...
The minimum number of rolls that should be made is 31.
The total number of golf balls is 60. To model this situation using a normal distribution the
number of dependent trials must be less than 10% of the population.
0.10(60) = 6
So, the maximum number of balls that could be selected is 5.
a) Technically, 10% of the population is 0.10(52), or 5.2. So, dealing five cards meets the
restriction.
13
b) The probability of drawing a heart, p, is , or 0.25.
52
μ np
5(0.25)
1.25
The mean is 1.25.
NP n
c) σ npq
NP 1
52 5
5 0.25 0.75
52 1
0.930
The standard deviation is about 0.930.
b) μ np σ npq
100(0.08) 100(0.08)(0.92)
8 2.713
For exactly 10 successes, calculate P(9.5 ≤ X ≤ 10.5) by inputting normalcdf(9.5,10.5,8,2.713).
The probability is about 0.1118.
b)
c) The height of each bar represents the probability of a particular number of heads.
e) Yes. Since the probability distribution is centred around a value and drops off symmetrically
to the right and left forming a bell-like shape, it is reasonable to model this experiment using a
normal distribtion.
f) μ np σ npq
12(0.5) 12(0.5)(0.5)
6 1.732
Use a graphing calculator and the function normalcdf to determine all P(x – 0.5 ≤ X ≤ x + 0.5),
where x = 0, 1, 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, μ = 6, and σ = 1.732
c) μ np σ npq
10(0.25) 10(0.25)(0.75)
2.5 1.369
The mean is 2.5, and the standard deviation is about 1.369.
e) The answers to parts a) and d) are very close. They both round to 5.8%.
0.9505
The probability that no ovens are dented is about 0.9505.
The normal approximation results in a higher probability of no dented ovens, 97.31% versus
95.05%.
b) Answers may vary. I chose to use the normal approximation for a hypergeometric distribution
because of the number of calculations needed to calculate P(X ≥ 10) using the hypergeometric
distribution.
b) No. A continuity correction is not needed because the population mean and standard
deviation are given for the normal distribution.
c) Answers may vary. No. With a probability of 26.5% that this scenario could happen by
chance, there is not enough information to conclude that the program was effective. More data
needs to be collected. If the program was effective, I would expect the probability of 4 or fewer
drivers being distracted to be higher.
a) The data are difficult to analyse in this form. It is not obvious whether the distribution is
uniform or not.
b) Use a table to determine the frequency for each interval. If all frequencies are equal, then the
distribution is uniform.
Percent Oxygen Frequency
31.7−31.8 3
31.8−31.9 3
31.9−32.0 3
32.0−32.1 3
32.1−32.2 3
This distribution is uniform.
a) Determine the height of the rectangle with base 2.8 and area 1.
bh A
3.8h 1
1
h
3.8
h 0.263
The height of the probability distribution is approximately 0.263.
b) The probability that an arrow has a length less than 71.1 cm equals the shaded area under the
graph from 69.2 cm to 71.1 cm.
P(69.2 x 71.1) area under graph
0.263(71.1 69.2)
0.4997
The probability that an arrow has a length less than 71.1 cm is 0.4997.
c) The probability that an arrow has a length between 70.6 cm and 71.6 cm equals the shaded
area under the graph from 70.6 cm to 71.6 cm.
P(70.6 x 71.6) area under graph
0.263(71.6 70.6)
0.263
The probability that an arrow has a length between 70.6 cm and 71.6 cm is 0.263.
a) and b)
b) To determine the probability, read the relative frequency from the table for 5–7 days. The
probability that a given fruit fly will die before the end of the week is 0.01.
c) To determine the probability, add the relative frequencies from 11 days to 17 days.
0.24 + 0.27 + 0.20 = 0.71
The probability that a given ball will launch at a speed of 59 km/h to 69 km/h is 0.71.
d) Answers may vary. No. While the distribution is somewhat bell-shaped, it is negatively
skewed.
Since the table does not contain a value for 0.000 001, use trial and error with the normalcdf
function on a graphing calculator. Use 5 for the lower bound, 99999 for the upper bound, 4.5 for
the mean, and test various values for the standard deviation.
Test Value for Normalcdf
Standard Deviation Result Comment
0.1 0.000 000 287 Too low.
0.11 0.000 002 744 Too high.
0.105 0.000 000 960 Very close.
0.1051 0.000 000 982 Almost.
0.1052 0.000 001 004 The standard deviation is about 0.1052 h.
a) A confidence level of 95% has a z-score of 1.960. Use the formula for the margin of error
with z = 1.960, p = 0.42, and n = 150.
p (1 p)
Ez
n
0.42(1 0.42)
1.960
150
0.079
The margin of error at a 95% confidence interval is about 7.9%.
b) The lower end of the range is 42% – 7.9%, or 34.1%. The upper end of the range is 42% +
7.9%, or 49.9%. The 95% confidence interval for the market share for the soft drink is 34.1% to
49.9%.
a) Use the margin of error formula for repeated samples. For a 90% confidence interval,
z = 1.645.
σ
Ez
n
9500
1.645
100
1562.75
The margin of error at a 90% confidence limit is 1562.75 km.
c) μ np σ npq
65(0.95) 65(0.95)(0.05)
61.75 1.757
The mean is 61.75, and the standard deviation is about 1.757.
a) Answers may vary. No. The class is not a representative sample of the population of the town
because it is only comprised of high school students.
b) The number of trials is less than 10% of the population. The normal approximation is
reasonable for this hypergeometric distribution.
NP n
c) μ np σ npq
NP 1
25 0.1
3500 25
2.5 25 0.1 0.9
3500 1
1.4948
The mean is 2.5, and the standard deviation is about 1.4948.
e) In this situation, n = 3500, r = 25, and a = 350. Use the indirect method to determine
P(x ≥ 5) = 1 – P(x < 5).
P( x 5) 1 P(0) P(1) P(2) P(3) P(4)
C C C C C C C C C C
1 350 0 3150 25 350 1 3150 24 350 2 3150 23 350 3 3150 22 350 4 3150 21
3500 C25 3500 C25 3500 C25 3500 C25 3500 C25
0.6430
Use the sum command on a graphing calculator.
The probability that at least five students have the flu is about 0.0973.
The normal approximation results in a higher probability, about 9.7% versus 9.0%.
Distance, mass, and rainfall are all continuous. The number of students with the flu in a given
class at your school would be expected to produce a discrete distribution. Answer C.
The number of customers, hamburgers, and watches are all discrete. The mass of a hawk recorded
during a migration would be expected to produce a continuous distribution. Answer B.
x
x
n
24.2 28.1 21.6 22.0 31.2
5
25.42
The mean is about 25.4 km/h. Answer C.
( x x ) 2
s
n 1
(24.2 25.4) 2 (28.1 25.4) 2 (21.6 25.4) 2 (22.0 25.4) 2 (31.2 25.4) 2
5 1
4.13
The standard deviation is about 4.13 km/h. Answer B.
The frequency associated with a jump between 180 cm and 190 cm is 120(0.117), or about 14.
Answer C.
The probability that a given flip-flop will have a length greater than 203 mm equals the shaded
area under the graph from 203 mm to 205 mm.
P(203 x 205) area under graph
0.1(205 203)
0.2
The probability that a given flip-flop will have a length greater than 203 mm is 0.2.
a) The number of students that lasted between 45 min and 50 min is 50(0.2), or 10.
a) and b)
c) Yes. Since the probability distribution is centred around a value and drops off mostly
symmetrically to the right and left forming a bell-like shape, the data appear to follow a normal
distribtion.
b) Work backwards. The percent of airliners were billed a nuisance fee is 0.4%.
The z-score associated with a probability of 1 – 0.004, or 0.996, is 2.65.
Determine the z-score for a noise level of 120 dB. Use x = 120, z = 2.65, and s = 6.7.
A confidence level of 99% has a z-score of 2.576. Use the formula for the margin of error with z
= 2.576, p = 0.84, and n = 500.
p (1 p)
Ez
n
0.84(1 0.84)
2.576
500
0.042
The margin of error at a 99% confidence interval is about 4.2%.
The lower end of the range is 84% – 4.2%, or 79.8%. The upper end of the range is 84% + 4.2%,
or 88.2%. The 99% confidence interval for the portion of parcels delivered within 30 min is
79.8% to 88.2%.
a) This is a hypergeometric probability distribution because there are two outcomes, success and
failure, and all trials are dependent.
b) The number of trials is less than 10% of the population. The normal approximation is
reasonable for this hypergeometric distribution.
NP n
c) μ np σ npq
NP 1
30 0.967
450 30
29.01 30 0.967 0.033
450 1
0.9463
The mean is 29.01, and the standard deviation is about 0.9463.