0% found this document useful (0 votes)
17 views55 pages

Continuous Probability Distributions Explained

Chapter 7 covers probability distributions for continuous variables, including prerequisite skills such as calculating areas, constructing frequency tables, and determining z-scores. It explains the concepts of continuous random variables, their properties, and how to compute probabilities associated with them. The chapter also includes examples and exercises to illustrate the application of these concepts in various scenarios.

Uploaded by

its.linh.c
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
17 views55 pages

Continuous Probability Distributions Explained

Chapter 7 covers probability distributions for continuous variables, including prerequisite skills such as calculating areas, constructing frequency tables, and determining z-scores. It explains the concepts of continuous random variables, their properties, and how to compute probabilities associated with them. The chapter also includes examples and exercises to illustrate the application of these concepts in various scenarios.

Uploaded by

its.linh.c
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

Chapter 7 Probability Distributions for Continuous Variables

Chapter 7 Prerequisite Skills

Chapter 7 Prerequisite Skills Question 1 Page 318

a) Substitute l = 4.2 and w = 1.9.


A = lw
= (4.2)(1.9)
= 7.98
The area is 7.98 m2.

a) Substitute a = 1.3, b = 2.4, and h = 1.8.


1
A  ( a  b)h
2
1
 (1.3  2.4)(1.8)
2
 3.33
The area is 3.33 cm2.

Chapter 7 Prerequisite Skills Question 2 Page 318

a) Construct a frequency table.


Number of Coins Frequency
1 2
2 2
3 5
4 3
5 2
6 3
7 1

b) Add a relative frequency column to the table from part a).


Number of Coins Frequency Relative Frequency
1 2 0.111
2 2 0.111
3 5 0.278
4 3 0.167
5 2 0.111
6 3 0.167
7 1 0.056

MHR  Data Management 12 Solutions 1


c) Use a graphing calculator. The mean number of coins is about 3.7778. The standard deviation
for the number of coins is about 1.7675.

d) Use the formula for sample z-score.


xx
z
s
2  3.7778

1.7675
 1.0058...
The z-score for a student with two coins in her pocket is about –1.0058.

Chapter 7 Prerequisite Skills Question 3 Page 318

a) Use the formula for sample z-score with each class.


Class A Class B
xx xx
z z
s s
x  74 x  72
z z
4 6
Set equal and solve for x.
x  74 x  72

4 6
6( x  74)  4( x  72)
6 x  444  4 x  288
2 x  156
x  78

2 MHR  Data Management 12 Solutions


A mark of 78 would result in the same z-score for each class.

x  74 x  72
b) No. The equation  has only one solution.
4 6

Chapter 7 Prerequisite Skills Question 4 Page 318

In a sample of 80 bills, you would expect 0.4(80) or 32 25-cent bills.

Chapter 7 Prerequisite Skills Question 5 Page 318

a) 5! = 5(4)(3)(2)(1) 3! 10! 7!
= 120 b) P  c) P2  d) C4 
(3  3)! (10  2)! 4!(7  4)!
3 3 10 7

3! 10  9  8! 7  6  5  4!
  
0! 8! 4!3!
3(2)(1)  10  9 765
 
1  90 3  2 1
6  35

Chapter 7 Prerequisite Skills Question 6 Page 319

Order does not matter. Frank can choose his 5 rides in 17C5, or 6188 ways.

Chapter 7 Prerequisite Skills Question 7 Page 319

a) Order matters. Wayne can arrange the 12 books in 12P12, or 479 001 600 ways.

b) Wayne can arrange the books in 3! × 3! × 3! × 3!, or 1296 ways, if they are clumped in
groups of 3 for each grade in increasing order from left to right.

Chapter 7 Prerequisite Skills Question 8 Page 319

There are a total of 15 gumballs with 5 green gumballs. The probability of Sam choosing green is
5 1
, or .
15 3

Chapter 7 Prerequisite Skills Question 9 Page 319

Order matters. The six student can be lined up in 6P6, or 720 ways. Only 1 of those ways is in
alphabetical order. So, the probability that they are lined up in alphabetical order by first name is
1
.
720

MHR  Data Management 12 Solutions 3


Chapter 7 Prerequisite Skills Question 10 Page 319

total value of all prizes


E( X )   price per ticket
number of tickets sold
100  75  25
 2
200
1 2
 1
The expected value per ticket is –$1.

Chapter 7 Prerequisite Skills Question 11 Page 319

These are independent trials. For a number greater than 9, you could roll a 10, 11, or 12 with the
3 1
12-sided die. So, the probability of rolling a number greater than 9 is , or . This is a
12 4
1 3
binomial distribution with p = and q = . The probability that a roll comes up greater than 9
4 4
4 3
1 3
exactly four times in seven rolls is 7 C4     , or about 0.0577.
4 4

Chapter 7 Prerequisite Skills Question 12 Page 319

12 1
These are independent trials. The probability of the barrier being down is , or . This is a
60 5
1 4
binomial distribution with p = and q = . The probability that exactly 20 cars will find the
5 5
20 80
1  4
barriers down in 100 cars is 100 C20     , or about 0.0993.
5  5

Chapter 7 Prerequisite Skills Question 13 Page 319

There are 26 students and 4 have red hair. Four are to be selected: n(S) = 26C4.
For the probability of 4 red heads, n(A) = 4C4 × 22C0.
C  C
P ( A)  4 4 22 0
26 C4

1

14 950
 6.689  105
The probability that all four have red hair is about .0067%.

Chapter 7 Prerequisite Skills Question 14 Page 319

Of the 1000 students at Eastdale Secondary School, 420 are boys and 580 are girls.
Ten are to be selected: n(S) = 1000C10.
For the probability of 5 boys and 5 girls, n(A) = 420C5 × 580C5.

4 MHR  Data Management 12 Solutions


C5  580 C5
P( A)  420

1000 C10

 0.2170
The probability that an equal number of boys and girls will receive a free sample is about 21.7%.

Chapter 7 Prerequisite Skills Question 15 Page 319

1
a) These are independent trials. The probability of a 50%-discount card is . This is a
100
1 99
binomial distribution with p = and q = . The probability that exactly 3 customers
100 100
3 247
 1   99 
received a 50%-discount card in 250 customers is 250 C3     , or about 0.2149.
 100   100 

b) Use the direct method to determine P(1 < x < 4).


P(1  x  4)  P(2)  P(3)
 250 C2 (0.01)2 (0.99) 248  250 C3 (0.01)3 (0.99) 247
 0.4750...
The probability that more than 1 but fewer than 4 customers received a 50%-discount card in 250
customers is about 47.5%.

Chapter 7 Section 1 Continuous Random Variables

Chapter 7 Section 1 Example 1 Your Turn Page 323


a) First, determine the height of the rectangle with base 4 and area 1.
bh  A
4h  1
1
h
4
h  0.25
The probability for a single value of a continuous distribution is 0. So, the probability that Chris
changes a tire in less than 9 min equals the shaded area under the graph from 8.0 min to 9.0 min.
In other words, it does not matter if the endpoint is included.
P(8.0  x  9)  area under graph
 0.25(9.0  8.0)
 0.25
The probability that Chris changes a tire in less than 9 min is 0.25.

b) The probability that a random variable assumes a value between a and b is the area under the
curve between a and b. The probability that it takes Chris between 9 min and 11.5 min to change
a tire equals the shaded area under the graph from 9.0 min to 11.5 min.
P(9  x  11.5)  area under graph
 0.25(11.5  9.0)
 0.625
The probability that it takes Chris between 9 min and 11.5 min to change a tire is 0.625.

MHR  Data Management 12 Solutions 5


c) If you pick a single value such as 10 min, the rectangle under the graph will have a width of 0
min. The probability for a single value of a continuous distribution is 0. The probability that it
takes exactly 10 min is 0.

Chapter 7 Section 1 Example 2 Your Turn Page 326

a) and b)

c) The mean cubit length in the class appears to be around 523 mm.

d) All outcomes are not equally likely. So, this is not a uniform distribution.

Chapter 7 Section 1 R1 Page 327

The number of guests that occupy the hotel each day is discrete data, while the time a guest waits
for an elevator is continuous data. It is possible to list all values of the discrete distribution, since
these would be values from 0 up to the maximum capacity of the hotel. It is not possible to list all
values of the continuous distribution, since time can be recorded in fractions of a second.

Chapter 7 Section 1 R2 Page 327

Answers may vary. The various height measurements could be caused by several things. For
example, Maya standing differently each time, bad measuring technique, and misreading the
measuring device. This is most likely measurement error, not bias.

Chapter 7 Section 1 R3 Page 327

Answers may vary. The probability that a variable falls within a range of values is equal to the
area under the probability density graph for that range of values. The area method cannot be used
for single values of a continuous variable, only for a range of values. A continuous random
variable can take on an infinite number of values. The probability that it will equal a specific
value is always zero.

Chapter 7 Section 1 Question 1 Page 327

a) Mass represents a continuous variable. This will not result in a discrete distribution.

6 MHR  Data Management 12 Solutions


b) The value of a card represents a discrete variable and would be expected to result in a discrete
distribution.

c) Barometric pressure represents a continuous variable. This will not result in a discrete
distribution.

Chapter 7 Section 1 Question 2 Page 327

a) The number of students with blue eyes represents a discrete variable and would not result in a
continuous distribution.
b) Weight represents a continuous variable and would be expected to result in a continuous
distribution.

c) The number of cartons of milk represents a discrete variable and would not result in a
continuous distribution.

d) The number of defective tablets represents a discrete variable and would not result in a
continuous distribution.

Chapter 7 Section 1 Question 3 Page 328

a) The distribution does not appear to be uniform because each waist size interval is not equally
likely.

b) The frequency associated with a waist size from 34 to 36 is 0.300(300), or 90.

c) If a data value falls on the boundary between two intervals, it is usually placed in the lower
interval. So, the additional waist size of 38 should go in the 36–38 interval.

Chapter 7 Section 1 Question 4 Page 328

a) The data are difficult to analyse in this form. It is not obvious whether the distribution is
uniform or not.

b) Use a table to determine the frequency for each interval. If all frequencies are equal, then the
distribution is uniform.
Volume of Sample (mL) Frequency
54−55 4
55−56 4
56−-57 4
57−58 4
58−59 4
59−60 4
60−61 4
61−62 4
62−63 4
63−64 4
64−65 4
This distribution is uniform.

MHR  Data Management 12 Solutions 7


Chapter 7 Section 1 Question 5 Page 328

a) Use a table to determine the frequency for each interval. If all frequencies are equal, then the
distribution is uniform.
Pressure (psi) Frequency
2995−2997 8
2997−2999 6
2999−3001 7
3001−3003 7
3003−3005 4
The distribution is not uniform.

b) In general, pressure is a continuous variable. So, the distribution is continuous.

c) Answers may vary. A possible reason for why there are no values with decimal places is that
the gauge used to measure the pressure only has a whole-number scale.

Chapter 7 Section 1 Question 6 Page 329

a) The probability of 30 min or less equals the shaded area under the graph from 25.0 min to
30.0 min.
P(25.0  x  30.0)  area under graph
 0.05(30.0  25.0)
 0.25
The probability a contestant will finish the triathlon in 30 min or less is 0.25.

b) The probability of 30 min to 40 min equals the shaded area under the graph from 30.0 min to
40.0 min.
P(30.0  x  40.0)  area under graph
 0.05(40.0  30.0)
 0.5
The probability a contestant will finish the triathlon in 30 min to 40 min is 0.5.

c) This type of distribution was used because all times are equally likely.

Chapter 7 Section 1 Question 7 Page 329

a) From the diagram, there are 88 keys including both black and white keys.

b) Answers may vary. I do not expect the distribution to be uniform, because I doubt that all
keys will be used in a piece of music.

c) Answers may vary. Since each key represents a specific note (frequency), the distribution in
part b) is discrete.

8 MHR  Data Management 12 Solutions


Chapter 7 Section 1 Question 8 Page 329

Answers may vary.


a) If a trombone player plays many notes at random frequencies, I would guess that he is
constantly moving the slider to new positions while blowing. Since the trombone can play all
frequencies between notes, I would expect the distribution to be continuous.

b) If a trombone player plays a musical composition, he plays specific notes. I would expect the
distribution to be discrete.

Chapter 7 Section 1 Question 9 Page 330

a) There are six possible lengths for the sardines. A quick inspection shows that there are six
occurrences of 99 mm. So, it is not possible for the 24 sardine sample to be a uniform
distribution.

b) Determine the height of the rectangle with base 5 and area 1.


bh  A
5h  1
1
h
5
h  0 .2
The manager should use 0.2 for the height of the probability density graph.

c)

d) The probability of a length less than or equal to 98 mm equals the shaded area under the
graph from 95.0 mm to 98.0 mm.
P(95.0  x  98.0)  area under graph
 0.2(98.0  95.0)
 0.6
The probability that a sardine has a length less than or equal to 98 mm is 0.6.

Chapter 7 Section 1 Question 10 Page 330

The area method cannot be used for single values of a continuous variable. So, Jon is incorrect.
The probability that a variable falls within a range of values is equal to the area under the
probability density graph for that range of values. The area shaded is from 2.5 to 3.5. So, Sunita is
correct.

MHR  Data Management 12 Solutions 9


Chapter 7 Section 1 Question 11 Page 330

a) The low and high cutoff values for this distribution is 5%, or 95 kΩ to 105 kΩ.

b) Determine the height of the rectangle with


base 10 and area 1.
bh  A
10 h  1
1
h
10
h  0.1

c) See part b). The height of the probability distribution is 0.1.

d) The low and high cutoff values for 0.25%, or 99.75 kΩ to 100.25 kΩ.
P(99.75  x  100.25)  area under graph
 0.1(100.25  99.75)
 0.05
The probability that a given resistor is within .25% of the stated resistance value is 0.05.

Chapter 7 Section 1 Question 12 Page 331

a) Speed is a continuous variable, so the data is continuous.

b) The data ranges from 64.6 to 80.2. So, I chose intervals of width two, starting at 64.5.

c)
Speed (km/h) Frequency
64.5–66.5 4
66.5–68.5 4
68.5–70.5 3
70.5–72.5 5
72.5–74.5 5
74.5–76.5 3
76.5–78.5 5
78.5–80.5 1

d)

10 MHR  Data Management 12 Solutions


e)

f) The mean appears to be around 72.5 km/h.

1
g) Answers may vary. Since only one speed was recorded over 80 km/h, my estimate is , or
30
about 0.033.

Chapter 7 Section 1 Question 13 Page 331

Answers may vary. I think that hair colour should be considered continuous, since there are
unlimited numbers of shades between blonde, red, brown, and black.

Chapter 7 Section 1 Question 14 Page 331

Answers may vary.


a) Three instruments that can play only discrete frequencies are electric piano, organ, and harp.

b) Three instruments that can play any frequency over a range are bass, cello, violin.

Chapter 7 Section 1 Question 15 Page 331

Answers may vary.


a) Use the formula for the area of a trapezoid with height 1 to calculate a probability to the left
or right of 0. Then, determine the sum if needed.

b) Calculate the area of three trapezoidal bars: 0 to 1, 1 to 2, and 2 to 3.


0 to 1: Substitute a = 0.4, b = 0.25, and h = 1.
1
A  ( a  b) h
2
1
 (0.4  0.25)(1)
2
 0.325
1 to 2: Substitute a = 0.25, b = 0.05, and h = 1.
1
A  (a  b)h
2
1
 (0.25  0.05)(1)
2
 0.15
2 to 3: Substitute a = 0.05, b = 0.0, and h = 1.

MHR  Data Management 12 Solutions 11


1
A  ( a  b) h
2
1
 (0.05  0.0)(1)
2
 0.025
The probability that the value of x lies between 0 and 3 is 0.325 + 0.15 + 0.025, or 0.5.

Chapter 7 Section 2 The Normal Distribution and z-Scores

Chapter 7 Section 2 Example 1 Your Turn Page 336

a) To determine the probability, read the relative frequency from the table for waist size 30–32
in. The probability that a customer has a waist size between 30 and 32 in. is 0.175.

b) To determine the probability, add the relative frequencies from 36 in. to 42 in.
0.16 + 0.045 + 0.000 = 0.205
The probability that a customer has a waist size of more than 36 in. is 0.205.

c) To determine the probability, add the relative frequencies from 30 in. to 36 in.
0.175 + 0.295 + 0.300 = 0.770
The probability that a customer has a waist size between 30 in. and 36 in. is 0.770.

d) Probabilities for a continuous distribution cover a range of values. The probability that a
waist size will be exactly 38 in. would require an interval width of 0 in., which is not available on
the table. The probability that a waist size will be exactly 38 in. is zero.

Chapter 7 Section 2 Example 2 Your Turn Page 340

a) Use a graphing calculator.

The mean, x , is 35.2 yd and the standard deviation, Sx, is


about 9.567 yd.

b) Choose –99999 as the lower bound and an upper bound of 30. Use the values for the mean
and standard deviation that were calculated in part a): normalcdf(–99999, 30, 35.2, 9.567).

The probability that Kunal kicks a distance of less than 30 yd is about 0.2934.

12 MHR  Data Management 12 Solutions


c) Choose 20 as the lower bound and an upper bound of 30. Use the values for the mean and
standard deviation that were calculated in parts a): normalcdf(20, 40, 35.2, 9.567).

The probability that Kunal kicks a distance of 20 yd to 40 yd is about 0.6360.

Chapter 7 Section 2 R1 Page 341

Answers may vary. As Kunal practises and gains more skill, I expect the mean and standard
deviation to change. As he improves and becomes more consistent, I expect the mean to increase
and the standard deviation to decrease.

Chapter 7 Section 2 R2 Page 341

The z-scores for a normal distribution follow a normal distribution themselves, with a mean of 0
and a standard deviation of 1. From the graph, P(z < –5) is located far in the left tail. So, the
probability that far from the central peak is essentially zero.

Chapter 7 Section 2 Question 1 Page 341

Use the formula for sample z-score with x = 13.8, x = 10.2, and s = 2.4.
xx
z
s
13.8  10.2

2.4
 1.5
The z-score for a trip that takes 13.8 min is 1.5. Answer C.

Chapter 7 Section 2 Question 2 Page 341

Use the formula for sample z-score with z = –0.5, x = 10.2, and s = 2.4.
xx
z
s
x  10.2
0.5 
2.4
0.5(2.4)  x  10.2
1.2  x  10.2
x9
The trip with a z-score of –0.5 took 9 min. Answer A.

MHR  Data Management 12 Solutions 13


Chapter 7 Section 2 Question 3 Page 341

A trip that is twice as long as the mean is 2(10.2), or 20.4 min. Recall that the probability that a
variable takes on a specific value is [Link] probability of a trip taking exactly 20.4 min is 0.

Chapter 7 Section 2 Question 4 Page 341

Answers may vary. Data collected one week could be different from another because of the path
chosen, traffic lights, or weather condition. Roberta could obtain more reliable values for the
mean and standard deviation by combining the data from the two weeks.

Chapter 7 Section 2 Question 5 Page 342

a)

b) The mean life of these light bulbs appears to be around 437.5 days.

c)
Lifetime (days) Frequency Relative Frequency
300–325 2 0.004
325–350 15 0.030
350–375 38 0.076
375–400 55 0.110
400–425 91 0.182
425–450 94 0.188
450–475 73 0.146
475–500 68 0.136
500–525 40 0.080
525–550 14 0.028
550–575 9 0.018
575–600 1 0.002

d) To determine the probability, add the relative frequencies from 300 days to 400 days.
0.004 + 0.030 + 0.076 + 0.110 = 0.220

14 MHR  Data Management 12 Solutions


The probability that a given light bulb will fail in 400 days or fewer is 0.220.

e) Answers may vary. Replace the light bulbs every 437.5 days, the estimated mean, to be
reasonably sure that there would never be a burned out light bulb.

Chapter 7 Section 2 Question 6 Page 342

a)
Speed of Ball
(km/h) Frequency
49–54 3
54–59 10
59–64 19
64–69 7
69–74 1

b)

c)

d) Answers may vary. While the distribution is centred around a central value, it does not drop
off symmetrically to the left and right.

e) The mean speed of the ball appears to be around 61.5 days.

MHR  Data Management 12 Solutions 15


f)
Speed of Ball
(km/h) Frequency Relative Frequency
49–54 3 0.075
54–59 10 0.250
59–64 19 0.475
64–69 7 0.175
69–74 1 0.025

g) To determine the probability, add the relative frequencies from 59 km/h to 69 km/h.
0.475 + 0.175 = 0.65
The probability that a given ball will launch at a speed of 59 km/h to 69 km/h is 0.65.

Chapter 7 Section 2 Question 7 Page 343

a)
Speed of Ball
(km/h) Frequency
49–51 2
51–53 1
53–55 1
55–57 4
57–59 5
59–61 8
61–63 10
63–65 3
65–67 2
67–69 3
69–71 1

b)

16 MHR  Data Management 12 Solutions


c)

d) The mean speed of the ball appears to be around 61 km/h.

e) Answers may vary. I think the smaller interval width made it easier to estimate the mean.

f) Answers may vary. The smaller interval width gives a clearer picture of the actual
distribution because it better approximates the shape of the frequency distribution.

Chapter 7 Section 2 Question 8 Page 343

a)
Horizontal Error
(cm) Frequency
(–10)–(–8) 1
(–8)–(–6) 3
(–6)–(–4) 6
(–4)–(–2) 6
(–2)–0 8
0–2 3
2–4 4
4–6 4
6–8 1

MHR  Data Management 12 Solutions 17


b)
Horizontal Error Relative
(cm) Frequency Frequency
(–10)–(–8) 1 0.0278
(–8)–(–6) 3 0.0833
(–6)–(–4) 6 0.1667
(–4)–(–2) 6 0.1667
(–2)–0 8 0.2222
0–2 3 0.0833
2–4 4 0.1111
4–6 4 0.1111
6–8 1 0.0278

c) To determine the probability, add the relative frequencies from –4 cm to 4 cm.


0.1667 + 0.2222 + 0.0833 + 0.1111 = 0.5833
The probability that the horizontal error is less than 4 cm to either side is 0.5833.

d) Answers may vary. A misadjusted sight with a bias to the left would result in more negative
values.

Chapter 7 Section 2 Question 9 Page 343

a) Use a graphing calculator.

The mean, x , is 35.145 cm.

b) See graphing calculator screen in part a). The standard deviation, Sx, is about 2.3612 cm.

c) Answers may vary. If a variable is expected to follow a normal distribution, you can take a
representative sample. However, there is not enough data to predict whether the distribution is
normal. The data given shows that distribution may be centred, but it is unclear whether it will
drop off symmetrically to the left and right.
Tree Height (cm) Frequency
30–32 3
32–34 2
34–36 7
36–38 7
38–40 1

18 MHR  Data Management 12 Solutions


Chapter 7 Section 2 Question 10 Page 344

a) Calculate the z-score for each student. Determine which student is performing at a higher
number of standard deviations from the respective means.

b) Answers may vary. The university must assume that the tests are comparable in material and
level of difficulty.

c)
Region A: Use x = 400, x = 350, and s = 35. Region B: Use x = 67, x = 62, and s = 5.
xx xx
z z
s s
400  350 67  62
 
35 5
 1.4286... 1

Since 1.429 > 1, the student from region A performed at a level that is further above average than
the student from region B.

Chapter 7 Section 2 Question 11 Page 344

Answers may vary.


a) marks on an exam, heights of plants, battery life, shoe size, masses of infants
Once data is collected and entered into technology, the remaining parts follow a similar procedure
to question 12.

Chapter 7 Section 2 Question 12 Page 344

a) Use a graphing calculator.

The mean, x , is 8.405.

b) See graphing calculator screen in part a). The standard deviation, Sx, is about 0.0196.

c) Lower limit: Use x = 8.38, x = 8.405, and s = 0.0196.


xx
z
s
8.38  8.405

0.0196
 1.2755...

Upper limit: Use x = 8.42, x = 8.405, and s = 0.0196.

MHR  Data Management 12 Solutions 19


xx
z
s
8.42  8.405

0.0196
 0.7653...
The z-scores that limit the acceptable range of the diameters are –1.28 and 0.77.

d) Use the table to determine the associated probability for a z-score of –1.28 as 0.1003 and the
associated probability for a z-score of 0.77 as 0.7794.
P(8.38 ≤ X ≤ 8.42) = P(X ≤ 8.42) − P(X ≤ 8.38)
= 0.7794 – 0.1003
= 0.6791
Then, the probability that a given manufactured piston falls outside the acceptable range is 1 –
0.6791, or 0.3209.

e) The number of pistons expected to be unacceptable is 500(0.3209), or about 160.

Chapter 7 Section 2 Question 13 Page 345

a) Calculate the z-score with x = 4000, x = 5200, and s = 1300.


xx
z
s
4000  5200

1300
 0.9230...
Use the table to determine the associated probability for a z-score of –0.92 as 0.1788. So, the
probability that the plant will release less than 4000 kg of uranium in a given year is 0.1788.

b) Calculate the z-score with x = 6000, x = 5200, and s = 1300.


xx
z
s
6000  5200

1300
 0.6153...
Use the table to determine the associated probability for a z-score of 0.62 as 0.7324. So, the
probability that the plant will release more than 6000 kg of uranium in a given year is 1 – 0.7324,
or 0.2676.

c) Within 10% of the mean is from 4680 to 5720.


Lower limit: Use x = 4680, x = 5200, and s = 1300.
xx
z
s
4680  5200

1300
 0.4

Upper limit: Use x = 5720, x = 5200, and s = 1300.

20 MHR  Data Management 12 Solutions


xx
z
s
5720  5200

1300
 0.4

Use the table to determine the associated probability for a z-score of –0.4 as 0.3446 and the
associated probability for a z-score of 0.4 as 0.6554.
P(4680 ≤ X ≤ 5720) = P(X ≤ 5720) − P(X ≤ 4680)
= 0.6554 – 0.3446
= 0.3108
Then, the probability that the release will be within 10% of the mean is 0.3108.

Chapter 7 Section 2 Question 14 Page 345

Answers may vary.


a) Marking on a curve has several interpretations and approaches: set the highest grade as 100%,
implement a flat-scale, set a bottom limit for fail, and use a bell curve.
To compare the two distributions, use the z-score. Determine the z-score for this year’s class and
then use that to determine the equivalent score using the historical mean and standard deviation.

b) This year’s mark: Use x = 70, x = 65, and s = 7.


xx
z
s
70  65

7
 0.7142...

Adjusted mark: Use z = 0.714, x = 72, and s = 5.


xx
z
s
x  72
0.714 
5
3.57  x  72
x  75.57
The adjusted mark would be 76.

c) Marks this year might be lower than expected because the calibre of the students has
declined.

Chapter 7 Section 2 Question 15 Page 345

a) Consider the formula for the mean of grouped data, which is the same as the sum of the
product of the midpoint of each interval, mi, and the relative frequency of each interval, rfi.

MHR  Data Management 12 Solutions 21


 f i mi
x
 fi
f1m1  f 2 m2  f n mn

 fi
f1m1 f 2 m2 f n mn
  
 fi  fi  fi
f1 f fn
 m1  m2 2  mn
 fi  fi  fi
 m1rf1  m2 rf 2  mn rf n
  mi rf i

b) Determine the midpoint of each interval.


f
Relative Frequency, rf =
Volume (mL) Frequency, f Midpoint, mi 200
490–492 0 491 0.000
492–494 0 493 0.000
494–496 2 495 0.010
496–498 11 497 0.055
498–500 43 499 0.215
500–502 81 501 0.405
502–504 48 503 0.240
504–506 14 505 0.070
506–508 1 507 0.005
508–510 0 509 0.000

x   mi rfi
 491(0)  493(0)  495(0.10)   509(0)
 501.08
The mean soft drink volume is 501.08 mL.

c) The value of the mean makes sense. It occurs in the centre of the bell-shaped frequency
polygon.

Chapter 7 Section 3 Applications of the Normal Distribution

Chapter 7 Section 3 Example Your Turn Page 348

a) Three standard deviations below the mean results in a z-score of −3, while three above the
mean results in a z-score of +3. Use the table to determine the probability that a value lies in the
range. Note: The table only contains z-score values from –2.99 to 2.99.
P(3  z  3)  P( z  2.99)  P( z  2.99)
 0.9986  0.0014
 0.9972

22 MHR  Data Management 12 Solutions


The probability that a value lies within three standard deviations of the mean is about 99.7%.

b) Use a graphing calculator and normalcdf(–3,3,0,1).

The probability that a value lies within three standard deviations of the mean is about 99.7%.

Chapter 7 Section 3 R1 Page 349

Answers may vary. Determine the z-score for a length of 4.7 cm.
Use x = 4.7, x = 5, and s = 0.1.
xx
z
s
4.7  5.0

0.1
 3
The probability that the length of a bolt is 4.7 cm or less is only about 0.1%, so this would be a
surprising value.

Chapter 7 Section 3 R2 Page 349

Answers may vary.


Each machine produces a normal distribution for the mass of honey in a jar. The distributions
have different means, one at 1.05 kg and the other at 1.2 kg. Each also produces a “tail” to the
left of 1.0 kg. If the standard deviations are correct, the two tail areas will be equal, i.e., 0.001.
Under these conditions, the results are possible.

MHR  Data Management 12 Solutions 23


Zoom in:

Chapter 7 Section 3 Question 1 Page 349

Since data collected from a large sample of people or naturally occurring phenomena usually
have a normal distribution, it is reasonable to expect the heights or students in a grade 3 class, the
mass of peanut butter in a sample of jars, and the distance that a person can throw a football
follow a normal distribution. Answer D.

Chapter 7 Section 3 Question 2 Page 349

True: The curve is symmetrical about a central peak.


False: The median is always less than the mean. The median and mean are equal.
False: All of the data values will occur within two standard deviations of the mean. 99.7% are
within three standard deviations of the mean.
False: The mean of a sample always matches the mean of the underlying normal distribution.
Answer A.

Chapter 7 Section 3 Question 3 Page 349

One cat out 40 represents 2.5%.


Determine the z-score for a mass of 5.2 kg.
Use x = 5.2, x = 4.2, and s = 0.5.
xx
z
s
5.2  4.2

0.5
2
Determine the probability that a value lies above this range.
1 – P(z ≤ 5.2) = 1 – 0.9772
= 0.0228
The probability that a mass is greater than two standard deviations is about 2.3%. Since this
represents approximately 1 cat, this claim makes sense.

24 MHR  Data Management 12 Solutions


Chapter 7 Section 3 Question 4 Page 350

Answers may vary. A quality control engineer would be interested in the mean and standard
deviation of the washers to ensure that they will actually fit onto the standard bolts. In particular,
the internal diameter of the washer must not be smaller the exterior diameter of the bolt shaft.

Chapter 7 Section 3 Question 5 Page 350

Answers may vary.


a) Use a graphing calculator with randNorm(700,13.2,10) and store them in L1.

b)

The mean, x , is about 701 mm and the standard deviation, Sx, is about 20.1 mm.

c) The random sample mean and standard deviation are fairly close to the underlying normal
distribution.

d)–f) Use a graphing calculator with randNorm(700,13.2,100) and randNorm(700,13.2,999).


Note the maximum number of trials is 999.
Sample Size Mean (mm) Standard Deviation (mm)
population μ = 700 σ = 13.2
10 x = 701 s = 20.1
100 x = 700 s = 14.9
1000 x = 700 s = 13.1

g) The larger the sample, the closer the sample measures are to the underlying normal
distribution.

MHR  Data Management 12 Solutions 25


Chapter 7 Section 3 Question 6 Page 350

a) Determine the z-score for a hand span of 16.4 cm.


Use x = 16.4, x = 21.8, and s = 2.4.
xx
z
s
16.4  21.8

2.4
 2.25
Use the table to determine the associated probability as 0.0122.
So, the probability that a student could not play an octave is 0.0122.

b) The distance needed to span one and one-half octaves is 24.6 cm.
Determine the z-score for a hand span of 24.6 cm.
Use x = 24.6, x = 21.8, and s = 2.4.
xx
z
s
24.6  21.8

2.4
 1.1666...
Use the table to determine the associated probability for a z-score of 1.17 as 0.8790.
Determine the probability that a value lies above this range.
1 – P(X < 24.6) = 1 – 0.8790
= 0.121
The probability that a student could play one and one-half octaves is 0.121.

Chapter 7 Section 3 Question 7 Page 350

a) Determine the z-score for 5 years.


Use x = 5, x = 8, and s = 1.5.
xx
z
s
58

1.5
 2
For a normal distribution, 95% of the data values lie within two standard deviations of the mean.
So of the remaining 5%, 2.5% will be below a z-score of –2.

b) Use the table to determine the associated probability for a z-score of –2 as 0.0228. Then, the
expected number of watches that the company will replace is 100 000(0.0228), or 2280. So, the
company should budget $5(2280), or $11 400 to replace watches under warranty.

26 MHR  Data Management 12 Solutions


c) Determine the z-score for 10 years.
Use x = 10, x = 8, and s = 1.5.
xx
z
s
10  8

1.5
 1.3333...
Use the table to determine the associated probability for a z-score of 1.33 as 0.9082. Since
90.82% of the data values are in this range, it is not a reasonable idea to offer a 10-year warranty.

Chapter 7 Section 3 Question 8 Page 350

a) A score of 60% is one standard deviation below the mean and a score of 80% is one standard
deviation above the mean. For a normal distribution, 68% of the data values fall within one
standard deviation of the mean. So, 200(0.68), or 136 students are expected to score between 60%
and 80%.

b) A score of 50% is two standard deviations below the mean and a score of 90% is two
standard deviations above the mean. For a normal distribution, 95% of the data values fall within
two standard deviations of the mean. So, 200(0.95), or 190 students are expected to score
between 50% and 90%.

c) Use the table to determine the associated probability for a z-score of –2 as 0.0228. So,
200(.0228), or about 5 students would be expected to score below 50%.

Chapter 7 Section 3 Question 9 Page 351

For a normal distribution, 95% of the data values fall within two standard deviations of the mean.
0.0001
So, the standard deviation needed to meet this requirement is , or 0.000 05 in.
2

Chapter 7 Section 3 Question 10 Page 351

a) Answers may vary. If the data set is reasonably large and the data fall into a symmetric bell
shape, then it seems reasonable to use a normal distribution to model the discrete data.

b) Answers may vary. I think a reasonable minimum number of hot dogs and buns is the mean,
120. This would cover 50% of the data in the analysis.

c) Determine the z-score for 25 hot dogs.


Use x = 25, x = 120, and s = 11.
xx
z
s
25  120

11
 8.6363...
The table does not contain a z-score for data outside three standard deviations from the mean. So,
the probability that fewer than 25 hot dogs are sold is basically 0.

MHR  Data Management 12 Solutions 27


d) Determine the z-score for 100 hot dogs and 140 hot dogs.
Lower limit: Use x = 100, x = 120, and s = 11.
xx
z
s
100  120

11
 1.8181...

Upper limit: Use x = 140, x = 120, and s = 11.


xx
z
s
140  120

11
 1.8181...
Use the table to determine the associated probability for a z-score of –1.82 as 0.0344 and the
associated probability for a z-score of 1.82 as 0.9656.
P(100 < X < 140) = P(X < 140) − P(X < 100)
= 0.9656 – 0.0344
= 0.9312
From May 1 to September 30 is 153 days. So, 153(.9312), or on about 142 days Frank should
expect to sell between 100 and 140 hot dogs.

e) Determine the z-score for 200 hot dogs.


Use x = 200, x = 120, and s = 11.
xx
z
s
200  120

11
 7.2727...
The table does not contain a z-score for data outside three standard deviations from the mean. So,
the probability that Frank would expect to sell 200 or more hot dogs is basically 0 and would not
happen on any day.

Chapter 7 Section 3 Question 11 Page 351

a) Use a graphing calculator.

The mean, x , is about 503 g and the standard deviation, Sx, is about 2 g.

28 MHR  Data Management 12 Solutions


b) The data in the table that actually fall within one standard deviation of the mean are in the
range of 501 g to 505 g. There are 24 values, or 80% or the data that actually fall within one
standard deviation of the mean.

c) The expected percent of the data within one standard deviation is 68%. So, these values are
more clustered around the mean.

Chapter 7 Section 3 Question 12 Page 351

a) Determine the standard deviation. From the table, 0.05 has a z-score of –1.65.
Use z = –1.65, x = 200, and x = 220.
xx
z
s
200  220
1.65 
s
1.65s  20
s  12.1212...
The standard deviation of the amount of corned beef is about 12.1 g.

b) To ensure that no more than 0.5% of the sandwiches contain less than 200 g of corned beef,
Rudy could buy a better slicing machine.
Determine the new standard deviation. From the table, 0.005 has a z-score of –2.58.
Use z = –2.58, x = 200, and x = 220.
xx
z
s
200  220
2.58 
s
2.58s  20
s  7.7519...
The standard deviation of the amount of corned beef is now about 7.8 g.

To ensure that no more than 0.5% of the sandwiches contain less than 200 g of corned beef, Rudy
could increase the slicing machine mean.
Determine the mean. From the table, 0.005 has a z-score of –2.58.
Use z = –2.58, x = 200, and s = 12.1.
xx
z
s
200  x
2.58 
12.1
31.218  200  x
x  231.218
The mean amount of corned beef is now about 231.2 g.

c) Answers may vary. Most likely increasing the slicing machine mean will be more cost
effective than buying a new machine.

MHR  Data Management 12 Solutions 29


Chapter 7 Section 3 Question 13 Page 351

a) For two standard deviations from the mean, k = 2. According to Chebyshev’s Theorem, no
1
more than 2 , or 25% of the values lie more than two standard deviations from the mean.
2

b) Since 95% of the data values lie within two standard deviations of a normal distribution, then
5% of the data lie outside this range. Chebyshev’s Theorem does not exactly agree, but could
since the 25% value is a maximum.

c) The proportion of values that must lie within k standard deviations of the mean is given by
1
1 2 .
k

Chapter 7 Section 4 Confidence Intervals

Chapter 7 Section 4 Example 1 Your Turn Page 355

a) A confidence level of 99% has a z-score of 2.576. Use the formula for the margin of error
with z = 2.576, p = 0.75, and n = 100.
p (1  p)
Ez
n
0.75(1  0.75)
 2.576
100
 0.112
The margin of error at a 99% confidence interval is about 11.2%.

b) Lower limit = 75% – 11.2% Upper limit = 75% + 11.2%


= 63.8% = 86.2%
The confidence interval is 63.8% to 86.2%.

c) Answers may vary. Hockey Night in Canada is watched by 75% of households. This estimate
is considered correct within 11.2%, 99 times out of 100.

Chapter 7 Section 4 Example 2 Your Turn Page 357

a) A confidence level of 90% has a z-score of 1.645. Use the formula for the margin of error
with z = 1.645, p = 0.13, and n = 400.
p(1  p)
Ez
n
0.13(1  0.13)
 1.645
400
 0.028
The margin of error at a 90% confidence interval is about 2.8%.

b) Use the rearranged form of the margin of error formula with z = 1.645, p = 0.13, and
E = 0.014.

30 MHR  Data Management 12 Solutions


z 2 ( p(1  p))
n
E2
1.6452 (0.13(1  0.13))

0.0142
 1561
To cut the margin of error in half, the sample size would have to increase to about 1561 pills.

Chapter 7 Section 4 Example 3 Your Turn Page 358

998  1234  1523  1760  937  1193  996  1002  986  1285  1163  1716
a) x
12
 1232.75

x 
n
420

12
 121.24
The mean of the sample is 1232.75 h, and the standard deviation of the sample is about 121.24 h.

b) For a 99% confidence interval, z = 2.576.


Lower limit  x  E Upper limit  x  E
 x  z x  x  z x
 1232.75  2.576(121.24)  1232.75  2.576(121.24)
 920.44  1545.06
The 99% confidence interval for the mean life of a light bulb is about 920.44 h to 1545.06 h.

Chapter 7 Section 4 R1 Page 359

No. The confidence level is the probability that a particular statistic is within the range indicated
by the margin of error. The confidence level’s related z-score is used to calculate the margin of
error.

Chapter 7 Section 4 R2 Page 359

The lower end of the range is 5.6% – 1.4%, or 4.2%. The upper end of the range is 5.6% + 1.4%,
or 7%. The 90% confidence interval for the mean defective tablets within one year is 4.2% to 7%.
This is the range of possible percents of defective tablets.
Answers may vary. This would help the manufacturer budget for returns and provide information
on the reliability of the manufacturing process.

Chapter 7 Section 4 R3 Page 359

A 90% confidence interval can be stated as 9 times out of 10. A 99% confidence interval can be
stated as 99 times out of 100.

MHR  Data Management 12 Solutions 31


Chapter 7 Section 4 Question 1 Page 359

a) For a 90% confidence interval, z = 1.645. The values will be about 1.6 standard deviations
from the mean.

b) For a 95% confidence interval, z = 1.960. The values will be about 2 standard deviations from
the mean.

c) For a 99% confidence interval, z = 2.576. The values will be about 2.6 standard deviations
from the mean.

Chapter 7 Section 4 Question 2 Page 359

17
The confidence level for this poll is , or 85%. Answer B.
20

Chapter 7 Section 4 Question 3 Page 359

True: The margin of error is 6.1%.


True: In a similar poll, 95% of the time between 50.9% and 63.1% of the people would be found
in favour.
True: The confidence interval is 57%  6.1%.
False: In a similar poll, 95% of the time 57% or more of the people would be found in favour. No
range of values is provided. Answer D.

Chapter 7 Section 4 Question 4 Page 359

σ
σx 
n
5000

100
 500
The expected standard deviation of the sample means is 500 km.

Chapter 7 Section 4 Question 5 Page 359

a) A confidence level of 99% has a z-score of 2.576. Use the formula for the margin of error
with z = 2.576, p = 0.82, and n = 25.
p (1  p)
Ez
n
0.82(1  0.82)
 2.576
25
 0.198
The margin of error at a 99% confidence interval is about 19.8%.

b) The lower end of the range is 82% – 19.8%, or 62.2%. The upper end of the range is 82% +
19.8%, or 101.8%. The 99% confidence interval for the exam marks is 62.2% to 101.8%.

32 MHR  Data Management 12 Solutions


c) Answers may vary. Students recorded a mean mark of 82%  19.8%, 99 times out of 100.

Chapter 7 Section 4 Question 6 Page 360

Answers may vary. For the original poll, the lower end of the range is 34% – 3.4%, or 30.6%.
The upper end of the range is 34% + 3.4%, or 37.4%. The 95% confidence interval for the
average percent of support is 30.6% to 37.4%. The other two polls are both outside of this range,
suggesting that support is not consistent.

Chapter 7 Section 4 Question 7 Page 360

48.9  50.1  47.5  50.1  47.1 σ


a) x σx 
20 n
 48.295 2.0

20
 0.447
The mean of the sample is about 48.3 g, and the standard deviation of the sample is about 0.45 g.
I assumed that the standard deviation stayed the same as the Single Crème line.

b) For a 95% confidence interval, z = 1.960.


Lower limit  x  E Upper limit  x  E
 x  zσ x  x  zσ x
 48.3  1.960(0.45)  48.3  1.960(0.45)
 47.4  49.2
The 95% confidence interval for the mean amount of filling is about 47.4 g h to 49.2 g.

c) Answers may vary. No. The company is not justified in claiming the line contains twice the
filling. The Double Crème line contains twice the filling when compared to approximately the
lower half of the Single Crème line distribution. So, for only less than half of the cookies is the
claim valid.

Chapter 7 Section 4 Question 8 Page 360

a) First, determine the standard deviation for the sample mean.


σ
σx 
n
8.5

25
 1.7
For a 95% confidence interval, z = 1.960.
Lower limit  x  E Upper limit  x  E
 x  zσ x  x  zσ x
 72.2  1.960(1.7)  72.2  1.960(1.7)
 68.868  75.532
The 95% confidence interval for the mean setting time of the concrete is about 68.9 min to 75.5
min.

MHR  Data Management 12 Solutions 33


b) Answers may vary. While using 72 minutes gives the most accurate description of the setting
time of the concrete, it is not as “nice” a number as, say 70 min. I would advise the manufacturer
that 70 min is a reasonable value for t, since it lies within the 95% confidence interval.

Chapter 7 Section 4 Question 9 Page 360

First, determine the standard deviation for the sample mean.


σ
σx 
n
2.5

50
 0.354
For a 95% confidence interval, z = 1.960.
Lower limit  x  E Upper limit  x  E
 x  zσ x  x  zσ x
 12  1.960(0.354)  12  1.960(0.354)
 11.3  12.7
The 95% confidence interval for the colour of honey is about 11 to 13.

Chapter 7 Section 4 Question 10 Page 360

For a confidence level of 99%, z = 2.576. Use the margin of error formula with E = 0.105 and σ =
0.05.
σ
Ez
n
σ2
E2  z2
n
σ2
n  z2 2
E
0.052
 2.5762
0.1052
 1.5047....
The number of businesses surveyed was about 2.

Chapter 7 Section 4 Question 11 Page 360

Use the margin of error formula with E = 0.2, n = 70, and σ = 1.


σ
Ez
n
n
zE
σ
70
 0.2
1
 1.67
According to the table, this z-score represents a confidence level of 0.9525 – 0.0475, or 90.5%.

34 MHR  Data Management 12 Solutions


Chapter 7 Section 4 Question 12 Page 361

For a confidence level of 95%, z = 1.960. Use the margin of error formula with E = 1.5 and σ =
5.9.
σ
Ez
n
σ2
E2  z2
n
σ2
n  z2 2
E
5.92
 1.9602
1.52
 59.4338....
Approximately 59 patients were in the study.

Chapter 7 Section 4 Question 13 Page 361

First, determine the standard deviation for the sample mean.


σ
σx 
n
2400

900
 80
For a 95% confidence interval, z = 1.960.
Lower limit  x  E Upper limit  x  E
 x  zσ x  x  zσ x
 16 000  1.960(80)  16 000  1.960(80)
 15 843.2  16156.8
The 95% confidence interval for the mean life of a hard drive is 15 843.2 h to 16 156.8 h.

Chapter 7 Section 4 Question 14 Page 361

a) A confidence level of 95% has a z-score of 1.960. Use the formula for the margin of error
with z = 1.960, p = 0.92, and n = 300.
p (1  p)
Ez
n
0.92(1  0.92)
 1.960
300
 0.031
The lower end of the range is 92% – 3.1%, or 88.9%. The upper end of the range is 92% + 3.1%,
or 95.1%. The 95% confidence interval for the load of tomatoes is 88.9% to 95.1%.

b) The lower end of the range is 0.889(41.992), or about 37.331. The upper end of the range is
0.951(41.992), or about 39.934. The 95% confidence interval for the mass of acceptable tomatoes
is 37.331 tonnes to 39.934 tonnes.

MHR  Data Management 12 Solutions 35


c) The lower end of the range is $94.40(37.331), or $3524.05. The upper end of the range is
$94.40(39.934), or $3769.77. The 95% confidence interval for the worth of acceptable tomatoes
is $3524.05 to $3769.77.

d) Now, the lower end of the range of acceptable tomatoes is 67%.

Use the formula for the margin of error with z = 1.960, p = 0.67, and n = 300.
p(1  p)
Ez
n
0.67(1  0.33)
 1.960
300
 0.053

The lower end of the range is 67%, or about 28.135. The upper end of the range is 67% + 5.3%,
or 72.3%. The interval for the mass of acceptable tomatoes is 28.135 tonnes to 30.360 tonnes.

The lower end of the range is $94.40(28.135), or $2655.91. The upper end of the range is
$94.40(30.360), or $2866. The interval for the worth of acceptable tomatoes is $2655.94 to
$2866.

Chapter 7 Section 4 Question 15 Page 361

The caption means that anywhere from 54% to 60% of students would vote for Adam and
anywhere from 48% to 54% of students would vote for Meghan with a 95% confidence level.

Chapter 7 Section 4 Question 16 Page 361

Answers may vary.


a) Since the masses of 20-year-old mean most likely follow a normal distribution, I would
expect to end up with more men close to the population mean, which is consistent with the bell-
shaped curve.

b) This results in the standard deviation of the sample mean being smaller than that of the
population.

σ
c) The effect in part b) fits with the formula for the standard deviation of the sample, σ x  .
n
The standard deviation of the population is divided by n , thus decreasing the value.

Chapter 7 Section 4 Question 17 Page 361

Use the margin of error formula with z = 1.960 and E ≤ 0.02.

36 MHR  Data Management 12 Solutions


p(1  p)
Ez
n
p(1  p)
0.02  1.960
n
0.02 p(1  p)

1.960 n
2
 0.02  p(1  p)
  
 1.960  n
2
 1.960 
n  p(1  p)  
 0.02 
n  9604 p(1  p)
Since the maximum value of the quadratic function f ( p)  9604 p(1  p) is 2401, the minimum
number of voters who must be interviewed is 2401.

Chapter 7 Section 5 Connections to Discrete Random Variables

Chapter 7 Section 5 Example 1 Your Turn Page 367

a) Use n = 25, p = 0.25, and q = 0.75.


np = 25 × 0.25 nq = 25 × 0.75
= 6.25 = 18.75
Since both of these are greater than 5, the normal approximation is reasonable in this case.

b) μ  np σ  npq
 25(0.25)  25(0.25)(0.75)
 6.25  2.165
The mean is 6.25, and the standard deviation is about 2.165.

c) If you want fewer than 8 successes, calculate P(X ≤ 7.5) by inputting


normalcdf(–999999,7.5,6.25,2.165). The probability is about 0.718.

MHR  Data Management 12 Solutions 37


Chapter 7 Section 5 Example 2 Your Turn Page 369

a) There are 90 socks in the drawer, of which 7 are chosen. The number of trials is less than
10% of the population. The normal approximation is reasonable for this hypergeometric
distribution.

b) There are 30 blue socks.


 NP  n 
μ  np σ  npq  
 NP  1 
 30 
 7   30  60  90  7 
 90   7    
 2.333  90  90  90  1 
 1.204
The mean is about 2.333, and the standard deviation is about 1.204.

c) For 3, 4, or 5 successes, calculate P(2.5 ≤ X ≤ 5.5 ) by inputting


normalcdf(2.5,5.5,2.333,1.204). The probability is about 0.4406.

Chapter 7 Section 5 R1 Page 369

Answers may vary. Any scenario that involves only two outcomes, success (p) and failure (q),
and the number of independent trials is large enough to meet the requirements of np > 5 and
nq > 5. For example, number of heads when flipping a coin or failure rate for quality control.

Chapter 7 Section 5 R2 Page 369

Answers may vary. Any scenario that involves two outcomes, success (p) and failure (q), in
dependent trials, and the number of dependent trials is less than 10% of the population.
Examples: choosing certain people to be on a committee or the number of a particular type of
card in a five-card hand.

38 MHR  Data Management 12 Solutions


Chapter 7 Section 5 R3 Page 369

Answers may vary. Using the approximation allows the probabilities of value ranges to be
calculated more easily than with the binomial or hypergeometric formulas.

Chapter 7 Section 5 R4 Page 369

Answers may vary. In the case of the normal approximation for a binomial distribution, any
scenario where np ≤ 5 or nq ≤ 5. In the case of the normal approximation for a hypergeometric
distribution, any scenario where the number of dependent trials is greater than or equal to 10% of
the population.

Chapter 7 Section 5 Question 1 Page 370

13
The probability of drawing a diamond, p, is , or 0.25.
52
μ  np σ  npq
 30(0.25)  30(0.25)(0.75)
 7.5  2.372
The mean is 7.5, and the standard deviation is about 2.372. Answer A.

Chapter 7 Section 5 Question 2 Page 370

There are 30 red jelly beans out of 200. 15 beans are drawn.
 NP  n 
σ  npq  
 NP  1 
 30  170  200  15 
 15    
 200  200  200  1 
 1.333
The standard deviation is about 1.333. Answer D.

Chapter 7 Section 5 Question 3 Page 370

6
The probability of a double, p, is . To model this situation using a normal distribution, np > 5
36
and nq > 5.
np  5 nq  5
 6   24 
n   5 n   5
 36   36 
 36   36 
n  5  n  5 
 6   24 
n  30 n  3.333...
The minimum number of rolls that should be made is 31.

MHR  Data Management 12 Solutions 39


Chapter 7 Section 5 Question 4 Page 370

The total number of golf balls is 60. To model this situation using a normal distribution the
number of dependent trials must be less than 10% of the population.
0.10(60) = 6
So, the maximum number of balls that could be selected is 5.

Chapter 7 Section 5 Question 5 Page 370

a) Technically, 10% of the population is 0.10(52), or 5.2. So, dealing five cards meets the
restriction.

13
b) The probability of drawing a heart, p, is , or 0.25.
52
μ  np
 5(0.25)
 1.25
The mean is 1.25.

 NP  n 
c) σ  npq  
 NP  1 
 52  5 
 5  0.25 0.75  
 52  1 
 0.930
The standard deviation is about 0.930.

Chapter 7 Section 5 Question 6 Page 370

a) In this case, n = 100, p = 0.08, and q = 0.92. Determine P(10).


P(10)  100 C10 (0.08)10 (0.92)90
 0.1024
The probability that exactly 10 of the cars contained one person is about 0.1024.

b) μ  np σ  npq
 100(0.08)  100(0.08)(0.92)
8  2.713
For exactly 10 successes, calculate P(9.5 ≤ X ≤ 10.5) by inputting normalcdf(9.5,10.5,8,2.713).
The probability is about 0.1118.

40 MHR  Data Management 12 Solutions


c) Answers may vary. Since both np and nq are greater than five, I would expect close
agreement between the two methods.

Chapter 7 Section 5 Question 7 Page 370

a) Use a graphing calculator to determine all P(x) = nCxpxqn−x , where n = 12, x = 0, 1, 2, 3, 4, 5,


6, 7, 8, 9, 10, 11, 12, p = 0.5, and q = 0.5.

b)

c) The height of each bar represents the probability of a particular number of heads.

d) The total area of the bars in the probability distribution is 1.

e) Yes. Since the probability distribution is centred around a value and drops off symmetrically
to the right and left forming a bell-like shape, it is reasonable to model this experiment using a
normal distribtion.

f) μ  np σ  npq
 12(0.5)  12(0.5)(0.5)
6  1.732
Use a graphing calculator and the function normalcdf to determine all P(x – 0.5 ≤ X ≤ x + 0.5),
where x = 0, 1, 2, 3, 4, 5, 6, 7, 8, 9, 10, 11, 12, μ = 6, and σ = 1.732

MHR  Data Management 12 Solutions 41


g) When tossing coins, the mean is  = 6 and the standard deviation is   (npq) .
 12  0.5  0.5
 1.732
Using normalcdf(–0.5,12.5,6,1.732), the probability is about 0.9998.

h) The areas are the same, 1.

Chapter 7 Section 5 Question 8 Page 370

a) In this case, n = 10, p = 0.25 and q = 0.75. Determine P(5).


P(5)  10 C5 (0.25)5 (0.75)5
 0.0584
The probability that exactly 5 diamonds are drawn is about 0.0584.

b) Check the restrictions.


np = 10 × 0.25 nq = 10 × 0.75
= 2.5 = 7.5
Since np is less than 5, it is not reasonable to model this distribution using a normal
approximation.

c) μ  np σ  npq
 10(0.25)  10(0.25)(0.75)
 2.5  1.369
The mean is 2.5, and the standard deviation is about 1.369.

d) For exactly 5 successes, calculate P(4.5 ≤ X ≤ 5.5) by inputting normalcdf(4.5,5.5,2.5,1.369).


The probability is about 0.0578.

e) The answers to parts a) and d) are very close. They both round to 5.8%.

Chapter 7 Section 5 Question 9 Page 371

a) In this situation, n = 200, r = 5, and a = 2. Determine P(0).


C  C
P(0)  2 0 198 5
200 C5

 0.9505
The probability that no ovens are dented is about 0.9505.

42 MHR  Data Management 12 Solutions


 NP  n 
b) μ  np σ  npq  
 NP  1 
 5  0.01
 200  5 
 0.05  5  0.01 0.99   
 200  1 
 0.2202
The mean is 0.05, and the standard deviation is about 0.2202.

c) For 0 successes, calculate P(–0.5 ≤ X ≤ 0.5) by inputting normalcdf(–0.5,0.5,0.05,0.2202).


The probability is about 0.9731.

The normal approximation results in a higher probability of no dented ovens, 97.31% versus
95.05%.

Chapter 7 Section 5 Question 10 Page 371

a) Use a hypergeometric distribution where n = 50, p = 0.12, q = 0.88, NP = 900.


 NP  n 
μ  np σ  npq  
 NP  1 
 50(0.12)
6  850 
 50(0.12)(0.88)  
 899 
 2.2343
The mean is 6, and the standard deviation is about 2.2343.
For 10 or more successes, calculate P(X ≥ 9.5) by inputting normalcdf(9.5,999999,6,2.2343). The
probability is about 0.0586.

b) Answers may vary. I chose to use the normal approximation for a hypergeometric distribution
because of the number of calculations needed to calculate P(X ≥ 10) using the hypergeometric
distribution.

MHR  Data Management 12 Solutions 43


Chapter 7 Section 5 Question 11 Page 371

a) Determine the z-score for 500 g.


Use x = 500, μ = 502.83, and σ = 1.95.
xμ
z
σ
500  502.83

1.95
 1.451...
Use the table to determine the associated probability for a z-score of –1.45 as 0.0735.
So, the probability that a honey jar contains less than 500 g of honey is 0.0735.

b) No. A continuity correction is not needed because the population mean and standard
deviation are given for the normal distribution.

c) A probability of 0.005 is represented by a z-score of –2.575. Use x = 500, z = –2.575, and


σ = 1.95.
xμ
z
σ
500  μ
2.575 
1.95
1.95(2.575)  500  μ
μ  500  1.95( 2.575)
 505.02125
The setting for the mean should be about 505.02 g.

Chapter 7 Section 5 Question 12 Page 371

a) Check the restrictions for n = 50 and p = 0.2.


np = 50 × 0.2 nq = 50 × 0.8
= 10 = 40
Since both of these are greater than 5, the normal approximation is reasonable in this case.
μ  np σ  npq
 50(0.2)  50(0.2)(0.8)
 10  2.828
For 30 or more successes, calculate P(X ≥ 29.5) by inputting normalcdf(29.5,999999,10,2.828).
The probability that Andre will pass the test is about 0.

44 MHR  Data Management 12 Solutions


b) Check the restrictions for n = 50 and p = 0.5.
np = 50 × 0.5 nq = 50 × 0.5
= 25 = 25
Since both of these are greater than 5, the normal approximation is reasonable in this case.
μ  np σ  npq
 50(0.5)  50(0.5)(0.5)
 25  3.536
For 30 or more successes, calculate P(X ≥ 29.5) by inputting normalcdf(29.5,999999,25,3.536).
The probability that Maria will pass the test is about 0.1016.

Chapter 7 Section 5 Question 13 Page 371

a) The expected number of drivers to be distracted is 0.05(120), or 6.

b) Check the restrictions for n = 120 and p = 0.05.


np = 120 × 0.05 nq = 120 × 0.95
=6 = 114
Since both of these are greater than 5, the normal approximation is reasonable in this case.
μ  np σ  npq
 120(0.05)  120(0.05)(0.95)
6  2.387
For 4 or fewer successes, calculate P(X ≤ 4.5) by inputting normalcdf(–999999,4.5,6,2.387). The
probability that 4 or fewer drivers would be distracted is about 0.2649.

c) Answers may vary. No. With a probability of 26.5% that this scenario could happen by
chance, there is not enough information to conclude that the program was effective. More data
needs to be collected. If the program was effective, I would expect the probability of 4 or fewer
drivers being distracted to be higher.

MHR  Data Management 12 Solutions 45


Chapter 7 Review

Chapter 7 Review Question 1 Page 372

a) The data are difficult to analyse in this form. It is not obvious whether the distribution is
uniform or not.

b) Use a table to determine the frequency for each interval. If all frequencies are equal, then the
distribution is uniform.
Percent Oxygen Frequency
31.7−31.8 3
31.8−31.9 3
31.9−32.0 3
32.0−32.1 3
32.1−32.2 3
This distribution is uniform.

Chapter 7 Review Question 2 Page 372

a) Determine the height of the rectangle with base 2.8 and area 1.
bh  A
3.8h  1
1
h
3.8
h  0.263
The height of the probability distribution is approximately 0.263.

b) The probability that an arrow has a length less than 71.1 cm equals the shaded area under the
graph from 69.2 cm to 71.1 cm.
P(69.2  x  71.1)  area under graph
 0.263(71.1  69.2)
 0.4997
The probability that an arrow has a length less than 71.1 cm is 0.4997.

c) The probability that an arrow has a length between 70.6 cm and 71.6 cm equals the shaded
area under the graph from 70.6 cm to 71.6 cm.
P(70.6  x  71.6)  area under graph
 0.263(71.6  70.6)
 0.263
The probability that an arrow has a length between 70.6 cm and 71.6 cm is 0.263.

46 MHR  Data Management 12 Solutions


Chapter 7 Review Question 3 Page 373

a) and b)

c) The mean life of the fruit flies appears to be 14 days.

Chapter 7 Review Question 4 Page 373

a) There are a total of 100 fruit flies.


Lifetime (days) Frequency Relative Frequency
5–7 1 0.01
7–9 3 0.03
9–11 13 0.13
11–13 24 0.24
13–15 27 0.27
15–17 20 0.20
17–19 9 0.09
19–21 3 0.03

b) To determine the probability, read the relative frequency from the table for 5–7 days. The
probability that a given fruit fly will die before the end of the week is 0.01.

c) To determine the probability, add the relative frequencies from 11 days to 17 days.
0.24 + 0.27 + 0.20 = 0.71
The probability that a given ball will launch at a speed of 59 km/h to 69 km/h is 0.71.

Chapter 7 Review Question 5 Page 373

a) Use a graphing calculator.

The mean, x , is 39.225 cm.

MHR  Data Management 12 Solutions 47


b) See graphing calculator screen in part a). The standard deviation, Sx, is about 5.8260 cm.

c) Answers may vary. Create a frequency table.


Soybean Height
(cm) Frequency
25–30 2
30–35 5
35–40 8
40–45 9
45–50 4

d) Answers may vary. No. While the distribution is somewhat bell-shaped, it is negatively
skewed.

Chapter 7 Review Question 6 Page 373

Calculate the z-score with x = 65 000, x = 62 000, and s = 2500.


xx
z
s
65 000  62 000

2500
 1.2
Use the table to determine the associated probability for a z-score of 0.62 as 0.8849. So, the
probability that a graduate will find a job with a starting salary of more than $65 000 is
1 – 0.8849, or 0.1151.

Chapter 7 Review Question 7 Page 373

a) Determine the z-score for a speed of 120 km/h.


Use x = 120, x = 105, and s = 7.
xx
z
s
120  105

7
 2.14
Determine the probability that a value lies above this range.
1 – P(z ≤ 2.14) = 1 – 0.9838
= 0.0162
The percent of drivers that will accumulate demerit points is about 1.6%.

48 MHR  Data Management 12 Solutions


b) Determine the z-score for 99 km/h and 101 km/h.
Lower limit: Use x = 99, x = 105, and s = 7.
xx
z
s
99  105

7
 0.8571...

Upper limit: Use x = 101, x = 105, and s = 7.


xx
z
s
101  105

7
 0.5714...
Use the table to determine the associated probability for a z-score of –0.86 as 0.1949 and the
associated probability for a z-score of –0.57 as 0.2843.
P(99 < X < 101) = P(X < 101) − P(X < 99)
= 0.2843 – 0.1949
= 0.0894
The probability that a given vehicle has a speed between 99 km/h and 101 km/h is 0.0894.

Chapter 7 Review Question 8 Page 374

Since the table does not contain a value for 0.000 001, use trial and error with the normalcdf
function on a graphing calculator. Use 5 for the lower bound, 99999 for the upper bound, 4.5 for
the mean, and test various values for the standard deviation.
Test Value for Normalcdf
Standard Deviation Result Comment
0.1 0.000 000 287 Too low.
0.11 0.000 002 744 Too high.
0.105 0.000 000 960 Very close.
0.1051 0.000 000 982 Almost.
0.1052 0.000 001 004 The standard deviation is about 0.1052 h.

Chapter 7 Review Question 9 Page 374

a) A confidence level of 95% has a z-score of 1.960. Use the formula for the margin of error
with z = 1.960, p = 0.42, and n = 150.
p (1  p)
Ez
n
0.42(1  0.42)
 1.960
150
 0.079
The margin of error at a 95% confidence interval is about 7.9%.

b) The lower end of the range is 42% – 7.9%, or 34.1%. The upper end of the range is 42% +
7.9%, or 49.9%. The 95% confidence interval for the market share for the soft drink is 34.1% to
49.9%.

MHR  Data Management 12 Solutions 49


Chapter 7 Review Question 10 Page 374

a) Use the margin of error formula for repeated samples. For a 90% confidence interval,
z = 1.645.
σ
Ez
n
9500
 1.645
100
 1562.75
The margin of error at a 90% confidence limit is 1562.75 km.

b) Lower limit  x  E Upper limit  x  E


 190 000  1562.75  190 000  1562.75
 188 437.25  191562.75
The 90% confidence interval for the mean lifetime is 188 437.25 km to 191 562.75 km.

Chapter 7 Review Question 11 Page 374

a) In this case, n = 65, p = 0.95 and q = 0.05. Determine P(X ≥ 60).


P( X  60)  P(60)  P(61)  P(62)  P(63)  P(64)  P(65)
 65 C60 (0.95)60 (0.05)5  65 C61 (0.95)61 (0.05) 4  65 C62 (0.95) 62 (0.05)3  65 C63 (0.95)63 (0.05)2
 65 C64 (0.95)64 (0.05)1  65 C65 (0.95)65 (0.05) 0
 0.8941

The probability that 60 or more buses arrive on time is about 0.8941.

b) Check the restrictions.


np = 65 × 0.95 nq = 65 × 0.05
= 61.75 = 3.25
Since nq is less than 5, it is not reasonable to model this distribution using a normal
approximation.

c) μ  np σ  npq
 65(0.95)  65(0.95)(0.05)
 61.75  1.757
The mean is 61.75, and the standard deviation is about 1.757.

d) For 60 or more successes, calculate P(59.5 ≤ X ≤ 65.5) by inputting


normalcdf(59.5,65.5,61.75,1.757). The probability is about 0.8834.

50 MHR  Data Management 12 Solutions


e) The answer to part a) is slightly higher than that of part d), approximately 89% compared to
88%.

Chapter 7 Review Question 12 Page 374

a) Answers may vary. No. The class is not a representative sample of the population of the town
because it is only comprised of high school students.

b) The number of trials is less than 10% of the population. The normal approximation is
reasonable for this hypergeometric distribution.

 NP  n 
c) μ  np σ  npq  
 NP  1 
 25  0.1
 3500  25 
 2.5  25  0.1 0.9   
 3500  1 
 1.4948
The mean is 2.5, and the standard deviation is about 1.4948.

d) For at least 5 successes, calculate P(4.5 ≤ X ≤ 25.5) by inputting


normalcdf(4.5,25.5,2.5,1.4948). The probability is about 0.0905.

e) In this situation, n = 3500, r = 25, and a = 350. Use the indirect method to determine
P(x ≥ 5) = 1 – P(x < 5).
P( x  5)  1  P(0)  P(1)  P(2)  P(3)  P(4)
C  C C  C C  C C  C C  C
 1  350 0 3150 25  350 1 3150 24  350 2 3150 23  350 3 3150 22  350 4 3150 21
3500 C25 3500 C25 3500 C25 3500 C25 3500 C25

 0.6430
Use the sum command on a graphing calculator.

The probability that at least five students have the flu is about 0.0973.
The normal approximation results in a higher probability, about 9.7% versus 9.0%.

MHR  Data Management 12 Solutions 51


Chapter 7 Test Yourself

Chapter 7 Test Yourself Question 1 Page 375

Distance, mass, and rainfall are all continuous. The number of students with the flu in a given
class at your school would be expected to produce a discrete distribution. Answer C.

Chapter 7 Test Yourself Question 2 Page 375

The number of customers, hamburgers, and watches are all discrete. The mass of a hawk recorded
during a migration would be expected to produce a continuous distribution. Answer B.

Chapter 7 Test Yourself Question 3 Page 375

x
x
n
24.2  28.1  21.6  22.0  31.2

5
 25.42
The mean is about 25.4 km/h. Answer C.

Chapter 7 Test Yourself Question 4 Page 375

( x  x ) 2
s
n 1
(24.2  25.4) 2  (28.1  25.4) 2  (21.6  25.4) 2  (22.0  25.4) 2  (31.2  25.4) 2

5 1
 4.13
The standard deviation is about 4.13 km/h. Answer B.

Chapter 7 Test Yourself Question 5 Page 375

The frequency associated with a jump between 180 cm and 190 cm is 120(0.117), or about 14.
Answer C.

Chapter 7 Test Yourself Question 6 Page 375

The curve is not skewed to the left or right.


The median is equal to the mean.
95% of the data occur within two standard deviations of the mean.
The mean of a sample is not always less than the mean of the underlying normal distribution.
Answer B.

52 MHR  Data Management 12 Solutions


Chapter 7 Test Yourself Question 7 Page 376

The probability that a given flip-flop will have a length greater than 203 mm equals the shaded
area under the graph from 203 mm to 205 mm.
P(203  x  205)  area under graph
 0.1(205  203)
 0.2
The probability that a given flip-flop will have a length greater than 203 mm is 0.2.

Chapter 7 Test Yourself Question 8 Page 376

a) The number of students that lasted between 45 min and 50 min is 50(0.2), or 10.

b) To determine the probability, add the probabilities from 30 min to 60 min.


0.04 + 0.04 + 0.10 + 0.20 + 0.24 + 0.16 = 0.78
The probability that a student lasted less than an hour is 0.78.

Chapter 7 Test Yourself Question 9 Page 376

Calculate the z-score with x = 14, x = 15, and s = 0.75.


xx
z
s
14  15

0.75
 1.33
Use the table to determine the associated probability for a z-score of −1.33 as 0.0918. So, the
probability that a hamburger will receive less than 14 mL of ketchup is 0.0918.

Chapter 7 Test Yourself Question 10 Page 376

The z-score associated with a probability of 0.002 is –2.88.


Use z = –2.88, s = 20, and x = 550.
xx
z
s
x  550
2.88 
20
2.88(20)  x  550
x  492.4
The blade should be replaced after 492.4 h of use.

Chapter 7 Test Yourself Question 11 Page 376

a) Check the restrictions.


np = 1000 × 0.75 nq = 1000 × 0.25
= 750 = 250
Since np and nq are both greater than 5, it is reasonable to model this distribution using a normal
approximation.

MHR  Data Management 12 Solutions 53


b) μ  np σ  npq
 1000(0.75)  1000(0.75)(0.25)
 750  13.693
The mean is 750, and the standard deviation is about 13.693.

Chapter 7 Test Yourself Question 12 Page 376

a) and b)

c) Yes. Since the probability distribution is centred around a value and drops off mostly
symmetrically to the right and left forming a bell-like shape, the data appear to follow a normal
distribtion.

Chapter 7 Test Yourself Question 13 Page 376

a) Determine the z-score for a noise level of 120 dB.


Use x = 120, x = 108, and s = 6.7.
xx
z
s
120  108

6.7
 1.79
Determine the probability that a value lies above this range.
1 – P(z ≤ 1.79) = 1 – 0.9633
= 0.0367
The percent of airliners that will be billed a nuisance fee is about 3.7%.

b) Work backwards. The percent of airliners were billed a nuisance fee is 0.4%.
The z-score associated with a probability of 1 – 0.004, or 0.996, is 2.65.
Determine the z-score for a noise level of 120 dB. Use x = 120, z = 2.65, and s = 6.7.

54 MHR  Data Management 12 Solutions


xx
z
s
120  x
2.65 
6.7
2.65(6.7)  120  x
x  120  2.65(6.7)
x  102.245
The new mean noise level is about 102 dB.

Chapter 7 Test Yourself Question 14 Page 377

A confidence level of 99% has a z-score of 2.576. Use the formula for the margin of error with z
= 2.576, p = 0.84, and n = 500.
p (1  p)
Ez
n
0.84(1  0.84)
 2.576
500
 0.042
The margin of error at a 99% confidence interval is about 4.2%.

The lower end of the range is 84% – 4.2%, or 79.8%. The upper end of the range is 84% + 4.2%,
or 88.2%. The 99% confidence interval for the portion of parcels delivered within 30 min is
79.8% to 88.2%.

Chapter 7 Test Yourself Question 15 Page 377

a) This is a hypergeometric probability distribution because there are two outcomes, success and
failure, and all trials are dependent.

b) The number of trials is less than 10% of the population. The normal approximation is
reasonable for this hypergeometric distribution.

 NP  n 
c) μ  np σ  npq  
 NP  1 
 30  0.967 
 450  30 
 29.01  30  0.967  0.033  
 450  1 
 0.9463
The mean is 29.01, and the standard deviation is about 0.9463.

d) For 27 or more successes, calculate P(26.5 ≤ X ≤ 30.5) by


inputting normalcdf(26.5,30.5,29.01,0.9463). The probability
is about 0.9383.

MHR  Data Management 12 Solutions 55

You might also like