0% found this document useful (0 votes)
330 views7 pages

Probability Distributions and Solutions

This document provides solutions to exercises involving probability distributions. It examines Bernoulli, binomial, hypergeometric, Poisson, and normal distributions. It calculates probabilities, means, variances, and compares exact to approximate distributions for various scenarios involving random variables like number of successes in a trial. Computer packages are used to calculate some probabilities denoted with an asterisk.

Uploaded by

pilas_nikola
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
330 views7 pages

Probability Distributions and Solutions

This document provides solutions to exercises involving probability distributions. It examines Bernoulli, binomial, hypergeometric, Poisson, and normal distributions. It calculates probabilities, means, variances, and compares exact to approximate distributions for various scenarios involving random variables like number of successes in a trial. Computer packages are used to calculate some probabilities denoted with an asterisk.

Uploaded by

pilas_nikola
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
  • Solution Exercise 9.1
  • Solution Exercise 9.6
  • Solution Exercise 9.5
  • Solution Exercise 9.9
  • Solution Exercise 9.7
  • Solution Exercise 9.8
  • Solution Exercise 9.10
  • Solution Exercise 9.11
  • Solution Exercise 9.12
  • Solution Exercise 9.13
  • Solution Exercise 9.16
  • Solution Exercise 9.15
  • Solution Exercise 9.14
  • Solution Exercise 9.17

Below, the indicator (*) is used to denote that a computer package is used for the

calculation of a probability.

Solution Exercise 9.1


a. Bernoulli distribution; p = 0.35; E (X ) = 0.35 and V (X ) = 0.2275
b. Binomial distribution; n = 100 and p = 0.35; E (Y ) = 35 and V (Y ) = 22.75
c. Hypergeometric distribution; n = 100, M = 350 and N = 1000. E (Y ) = 35 and
900
V (Y ) = 22.75 = 20.4955
999

Solution Exercise 9.2


a. E (Y ) = 6 and V (Y ) = 5.4

 60 
b. P(Y  5) =    0.15  0.955 = 0.1662; (*)
5 
 60 
P(Y  6) =    0.16  0.954 = 0.1693; answer: 0.3355 (*)
6 
c. 0.2710 and 1 – P(Y  6) = 1 – 0.6065 = 0.3935
d. The answer of b. and the two answers of c. add up to 1, as it should.

Solution Exercise 9.3


a. Y  H(5; 115, 200)
b. E (Y ) = 5115/200 = 2.8750
V (Y ) = 5  (115/200)  (85/200)  (195/199) = 1.1973

115  85 
  
 3  2 
c. P(Y  3) = = 0.3476
 200 
 
5 
d. P(Y  2) = 0.2553; P(Y  2) = 0.0129 + 0.0918 = 0.1047 (*)
e. Bin(5, 0.575)
P(Y  3) = 0.3434; P(Y  2) = 0.2538; P(Y  2) = 0.1077 (*)

Solution Exercise 9.4


a. E (Y ) = 5.8 and V (Y ) = 5.8

1
5.85 5.8
b. P(Y  5) = e = 0.1656
5!
c. P(Y  6) = 0.1601; P(Y  6) = 1 – P(Y  6) = 1 – 0.6384 = 0.3616 (*)

Solution Exercise 9.5


a. Since the sample is drawn from a large population, it can be considered as
drawn with replacement. That is why the 100 trials (the randomly drawings of
the apples and qualifying them) can be considered to be independent and
identical. Since interest is in the number of ‘good’ apples in the sample, the
success probability p is the proportion of the ‘good’ apples in the whole lot.
The random experiment is binomial.
b. The sample of TVs is obviously drawn without replacement (from a small
population of TVs). Hence, the trials are NOT independent and the experiment
is not binomial.
c. The 20 subsequent throws of the die are done independently, with the same
die. The 20 trials are independent and identical; the experiment is binomial.
The success probability is the probability of throwing six eyes in one throw.
d. The weather on two consecutive days can hardly be considered as
independent. Not binomial!
e. The prices of a stock cannot be considered as independent. Not binomial!
f. Whether the daily returns of a stock (with respect to the day before) are
independent, is a matter of discussion. As a model, it can be assumed that the
returns on 15 consecutive days are independent repetitions of an experiment.
In this model, the combined experiment is binomial and p is the overall
proportion of positive returns.

Solution Exercise 9.6


a.
y 0 1
f(y) 0.98 0.02
Y  Bern(0.02)

F ( y) = 0 if y < 0
= 0.98 if 0  y < 1
=1 if y  1
b. E (Y ) = 0.02 and V (Y ) = 0.020.98 = 0.0196
c. Family of binomial distributions. Parameters: n = 100 and p = 0.02
d. E (Y ) = np = 2 and V (Y ) = np(1  p) = 1.96

2
100  100 
e. P(Y  0) =  0.98100 = 0.1326; P(Y  1) =  0.0210.9899 = 0.2707;
0  1 
100 
P(Y  2) =  0.0220.9898 = 0.2734;
2 
100 
P(Y  3) =  0.0230.9897 = 0.1823;
3 
100 
P(Y  4) =  0.0240.9896 = 0.0902
4 
Hence: P(Y  2) = 0.1326 + 0.2707 = 0.4033
P(Y  3) = 0.4033 + 0.2734 + 0.1823 = 0.8590
P(2  Y  5) = 0.1823 + 0.0902 = 0.2725

Solution Exercise 9.7


Y = ‘number of managers with that opinion among next n participants’
 4
a. Here, n = 4. P(Y  1) =    0.341  0.663 = 0.3910
1 
8 
b. n = 8; P(Y  2) =    0.342  0.666 = 0.2675
 2
12 
c. n = 12; P(Y  3) =    0.343  0.669 = 0.2055
3 
 40 
d. n = 40; P(Y  10) =    0.3410  0.6630 = 0.0675
10 

Solution Exercise 9.8


a. X = ‘number of Windows Vista O/S’s among 100 computers’
E (X ) = 1000.0452 = 4.52; SDX ) = 100  0.0452  0.9548 = 2.0774
b. P( X  3) = 0.0098 + 0.0464 + 0.1087 + 0.1681 = 0.3330
So, P( X  4) = 1 – 0.3330 = 0.6670
c. Y = ‘number of Windows Vista O/S’s among 1000 computers’
P(Y  40) = 0.2407 (*)

Solution Exercise 9.9


a. X has a binomial distribution. Parameters: n = 5 and p = 1/3.

3
b. E (X ) = 5/3; V (X ) = 51/32/3 = 10/9; SD(X ) = 1.0541.

5 
c. P( X  k ) =   (1 / 3)k (2 / 3)5 k for k = 0, 1,   , 5. Hence:
k 

k 0 1 2 3 4 5
f(k) 0.1317 0.3292 0.3292 0.1646 0.0412 0.0041

d. P( X  3) = 0.1646 + 0.0412 + 0.0041 = 0.2099.

Solution Exercise 9.10


a. Experiment: randomly choosing a first year university student. The random
variable X that takes the value 1 if this student is female (and 0 otherwise) has
the distribution Alt (0.51) .
b. E (X ) = 0.51 and V (X ) = 0.510.49 = 0.2499.
c. Although the exact distribution is H(100; 20400, 40000), the ratio n / N is
very small and hence Bin(100, 0.51) is a very good approximation.
d. E (Y ) = 1000.51 = 51, V (Y ) = 1000.510.49 = 24.9900 and SD(Y ) =
4.9990.
e. P(Y  48) = P(Y  47) = 0.2419 (*);
P(Y  52) = 1  P(Y  52) = 1 – 0.6176 = 0.3824 (*);
P(45  Y  55) = P(Y  54)  P(Y  45) = 0.7579 – 0.1356 = 0.6223 (*)

Solution Exercise 9.11


a. Since the sample is randomly drawn and without replacement from a relatively
small population, the pdf of Y belongs to the family of hypergeometric
distributions. The parameters are n = 20, M = 40 and N = 200. Notation: Y 
H (20;40,200) .
b. E (Y ) = 20(40/200) = 4, V (Y ) = 20(40/200)(1-40/200)(180/199) =
2.8945; SD(Y ) = 1.7013.
c. Bin (20,0.2) .
d. Before the computer can be used, the probabilities have to be rewritten. For
instance:
P(Y  3) = 1  P(Y  3) = 1 – ( P(Y  0) +    + P(Y  3) )
P(2  Y  6) = P(3  Y  5) = P(Y  3) + P(Y  4) + P(Y  5)

4
Here are the answers: exact 0.1095 0.9243 0.5978 0.6232
approximated 0.1091 0.9133 0.5886 0.5981

Solution Exercise 9.12


a. X has the binomial distribution with parameters n = 100 and p = 0.05.
b. If necessary, see Appendix A1.9.

0.2

0.15

0.1

0.05

0
0 7 14 21 28 35 42 49 56 63 70 77 84 91 98

c. P( X  4) = 0.4360 (*)
d. P( X  3) = 1  P( X  3) = 1  P( X  2) = 1 – 0.1183 = 0.8817 (*)
e. P(1  X  4) = P( X  4)  P( X  0) = 0.4360 – 0.0059 = 0.4301 (*)
f. E (X ) = 1000.05 = 5 and V (X ) = 1000.050.95 = 4.75.

g. E (Pˆ ) = 0.05, V (Pˆ ) = V (X ) / (100)2 = 0.000475 and SD(Pˆ ) = 0.0218.


h. Possible outcomes: 0, 0.01, 0.02,   , 0.99, 1

P( Pˆ  0.075) = 1  P( Pˆ  0.075) = 1  P( X  7.5)

= 1  P( X  7) = 1  0.8720 = 0.1280 (*)

P(0.03  Pˆ  0.081) = P( Pˆ  0.081)  P( Pˆ  0.03)

= P( X  8)  P( X  3) = 0.9369 – 0.2578 = 0.6791 (*)

Solution Exercise 9.13


a. H(n = 10, M = 40, N = 100)
b. Bin (n  10, p  0.4)
c. In a.:  = 100.4 = 4;
2 = 100.40.6((100  10)/(100  1)) = 2.1818

5
In b.:  = 100.4 = 4; 2 = 100.40.6 = 2.4
d. With H(10; 40, 100):
 40  60 
  
P( X  2) =    = 0.1153
2 8
100 
 
10 
P( X  3) = 1 – P( X  0)  P( X  1)  P( X  2)
= 1 – 0.0044 – 0.0342 – 0.1153 = 1  0.1539 = 0.8461
P( X  2) = P( X  3) + P( X  2) = 0.8461 + 0.1153 = 0.9614

With Bin(10, 0.4):


P( X  2) = 0.1209
P( X  3) = 1 – 0.0060 – 0.0403 – 0.1209 = 0.8328
P( X  2) = 0.8328 + 0.1209 = 0.9537

Solution Exercise 9.14


a. Po(14)
72 7
b. X = ‘# cars in one hour’; P( X  2) = e = 0.0223
2!
c. Y = ‘# cars in three hours’; Y  Po(21); P(Y  15) = 0.1111 (*)

Solution Exercise 9.15


a. Since p = 0.0071 and n = 10000 it follows that the expectation is 71 and the
standard deviation is 8.3962. So:   3 = 45.8114 and   3 = 96.1886
Cheb.: P(46  Y  97)  0.888

7170  71 7171  71
b. P(Y  70) = e = 0.0473; P(Y  71) = e = 0.0473;
70! 71!
P(Y  72) = 0.0466
c. P(Y  71) = 0.5315 (if necessary, see Appendix A1.9) (*)
d. P(46  Y  97) = P(Y  97)  P(Y  45) = 0.9986 – 0.0006 = 0.9980 (*)

Solution Exercise 9.16


a. Po(15)
b. 15 and 15

6
c. P( X  14) = 1 – P( X  14) = 1 – 0.6694 = 0.3306 (*)
d. P(10  X  14) = P( X  11)  P( X  12)  P( X  13)
= 0.0663 + 0.0829 = 0.0956 = 0.2448

Solution Exercise 9.17

0.25
binomial
0.2 Poisson
hypergeom
0.15

0.1

0.05

0
0 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15

Common questions

Powered by AI

The binomial distribution is useful as a model for real-world processes where there are fixed numbers of trials, each with two possible outcomes (success or failure), such as quality control testing, clinical trials, and election predictions . It requires independence between trials and a constant probability of success across the trials. For example, in manufacturing, modeling the number of defective products allows for planning and quality assurance interventions. However, these applications are constrained when dependencies or varying probabilities are present, thus requiring consideration or alternative models like the hypergeometric or negative binomial distribution . Understanding these constraints ensures a reliable application of the binomial model across various domains.

To calculate the probability of exactly k successes in a hypergeometric distribution, use the hypergeometric probability formula: P(Y=k) = [C(M, k) × C(N-M, n-k)] / C(N, n), where C(a, b) represents combinations of a taken b at a time, M is the number of successes in the population, N is the total population size, and n is the sample size . This formula calculates the probability by considering the combinations of selecting k from the successes and n-k from the failures, divided by the total sample combinations from the population.

The binomial distribution is often used as an approximation for hypergeometric distributions when the sample size n is small relative to the population N, making the ratio n/N negligible . This approximation is valid because under these conditions, the probability of success remains nearly constant regardless of the outcome of previous draws (approximating independence), allowing the hypergeometric distribution's behavior to resemble that of the binomial distribution . The larger the population, the more accurate the approximation because the dependence introduced by sampling without replacement becomes minimal.

For a model using stock prices to be appropriately described by a binomial distribution, the daily returns must be independent repetitions of the experiment . This assumes that the daily returns are independent and identically distributed, each with a probability p of a positive return and 1-p for a negative return. However, this assumption is contentious, as stock prices can be influenced by numerous factors leading to dependency in returns. The validity of modeling stock prices binomially rests on the assumption that these dependencies are negligible or accounted for in the modeling . Thus, discerning the independence and consistency of stock price changes across observation periods is crucial.

The expected value and variance are fundamental in determining the probability of outcomes in a statistical experiment. The expected value, or mean, indicates the central tendency or average number of successes in experiments under repetition. Variance quantifies the spread or dispersion of the probability distribution, indicating how much the actual results may vary from the expected value . Together, they define the confidence and reliability of different outcomes occurring by illustrating the distribution's central spine and variability around it, enabling analysts to assess odds and make informed predictions based on distribution patterns.

The concept of independence is central to determining whether a statistical experiment is binomial. For an experiment to be binomial, each trial must be independent, meaning the outcome of one trial does not affect the outcome of another. For example, tossing a die multiple times is binomial because each toss is independent. Conversely, drawing samples without replacement from a small population, such as TVs, is not binomial because the samples affect each other's outcomes, violating the independence requirement . Independence ensures that trials do not influence each other, which is fundamental for modeling using a binomial distribution.

The variance of a binomial distribution with parameters n = 100 and p = 0.35 is calculated as V(Y) = np(1-p) = 100 × 0.35 × (1-0.35) = 22.75 . For a hypergeometric distribution with n = 100, M = 350, and N = 1000, the variance is V(Y) = np(1-p) × (N-n)/(N-1) = 100 × (350/1000) × ((1000-350)/1000) × 900/999 = 20.4955 . The variance of the binomial distribution is larger than that of the hypergeometric distribution due to the correction factor (N-n)/(N-1) in the hypergeometric distribution that reduces the variance.

Both hypergeometric and binomial distributions have the same expectation formula E(Y) = np because they both model scenarios involving a fixed number of successes in a series of trials . However, they differ in their variance calculations due to the nature of sampling. The binomial distribution assumes independence between trials, thus its variance is V(Y) = np(1-p). However, in a hypergeometric distribution, the lack of replacement in sampling introduces dependency between trials, which is adjusted with a finite population correction factor, resulting in V(Y) = np(1-p) × (N-n)/(N-1). This accounts for the reduction in variability because sampling without replacement ties the outcomes together.

The variance of a hypergeometric distribution includes a finite population correction factor calculated as (N-n)/(N-1), which adjusts the variance downwards . This adjustment is necessary because hypergeometric sampling is done without replacement, leading to dependence among trials. The correction factor accounts for the decreasing population size after each draw, which ties the outcomes more closely together. This contrasts with the binomial variance which does not require such an adjustment as the trials are independent, due to sampling with replacement or the assumption thereof . The correction ensures the variance accurately reflects the reduced variability due to sampling dependencies.

A Poisson distribution is used as an approximation for a binomial distribution when the number of trials n is large, and the probability of success p is small such that np is moderate (often approaching values around 10 or less). Under these circumstances, the binomial distribution’s skewness and shape converge to that of a Poisson distribution as the probability of any individual event occurring becomes low, but collectively these rare events produced an expected result . This approximation is particularly useful for simplifying complex binomial calculations in large-sample low-probability scenarios.

1 
Below, the indicator (*) is used to denote that a computer package is used for the 
calculation of a probability. 
 
Sol
2 
b. 
)
5
(

Y
P
 = 
8.5
5
!5
8.5

e
 = 0.1656 
c. 
)
6
(

Y
P
 = 0.1601; 
)
6
(

Y
P
 = 1 – 
)
6
(

Y
P
 = 1 – 0.63
3 
e. 
)
0
(

Y
P
 = 
100
98
.0
0
100






 = 0.1326; 
)1
(

Y
P
 = 
99
1 98
.0
02
.0
1
100






 = 0.2
4 
b. 
)
(X
E
 = 5/3; 
)
(X
V
 = 51/32/3 = 10/9; 
)
(X
SD
 = 1.0541. 
c. 
)
(
k
X
P

 = 






k
5
k
k

5)3
/
2
5 
Here are the answers: 
 
 
Solution Exercise 9.12 
a. X has the binomial distribution with parameters n = 100 and p = 0.
6 
In b.:  = 100.4 = 4; 2 = 100.40.6 = 2.4 
d. With H(10; 40, 100):  
)
2
(

X
P
 =  













7 
c. 
)
14
(

X
P
 = 1 – 
)
14
(

X
P
 = 1 – 0.6694 = 0.3306 (*) 
d. 
)
14
10
(

X
P
 
= 
)
13
(
)
12
(
)
11
(





You might also like