National Open University of Nigeria: School of Science and Technology
National Open University of Nigeria: School of Science and Technology
1
Course Code STT 211
2
PROBABILITY DISTRIBUTION I
MODULE 1
UNIT 1
PREREQUISITES
The main prerequisites for understanding the content of this book is a knowledge of
elementary algebra, set theory, mathematical induction, differentiation and integration.
Set
A set is any well-defined list or collection of objectives. The objectives comprising the
set are called its elements or members.
A set will be denoted by capital letters or symbols such as X,Y,A,B,….. and its elements
will be denoted by lower case letters x, y, a,b,…..
Example 1.1 Toss a cubical die once. There are six possible numbers that can appear. We
can write
Ω = {1, 2, 3, 4,5, 6}
Where Ω is a set consisting of six elements 1, 2, 3, 4, 5, 6 called the element of the set Ω.
There are essentially two ways of specify a particular set. One way if possible is by
listing its elements as in example1.1 above. The other way is by stating properties which
characterize the elements in the set. The above set Ω can be written as:
Ω – {x: is an integer, 1 < x < 6}
If x is an element Ω, the notation x € Ω means that x belongs to Ω. The negation of this
assertion i.e. the statement that x does not belong to Ω will be denoted by
x€Ω
thus, for the above example 2 € Ω but 8 € Ω.
Definition 1.2
Subset: A set A is a subset of a set if each element in A also belongs to Ω.
In example 1.1 above, the set
A ={1,3,5}= (x: x €Ω and x is odd}
Is a subset of Ω, that is each element of A is in Ω.
Two sets A and B are called equal if and only if they contain exactly the same elements.
3
Throughout this book, whenever the word set is used, itwill be interpreted to mean a
subset of a given set denoted by Ω. The set which contains no elements is called the Null
set.
Definition
Let A be a set. The elements which are not included in A also constitute a subset. This is
known as the complement of A and is denoted by Ac.
In example 1,1 if A {1,3,5}
Ac = {2,4,6} = {x: x Ω, x even}.
Definition:
Two sets A, B define two related sets. One of these is the set of all elements which
belongs to both sets A and B. this is called the intersection of A and B and is denoted by
A∩B. the other is the set of all the elements which occur in either A or B or both. This is
called the union of A and B denoted by A B.
Example 1.2
Let Ω = {1,2,3,4,5,6}, A = {1,3,5} and B = {2,3,5}
Then
A∩B = {3,5}
A ∩B = {1,2,3,5}
A and B contain 3 element each while A ∩ B contains 2 elements and A ∩B contains 4
laments.
Note that the number of elements in A ∩ B is not the sum of the number of elements in A
and B.
Definition: Difference of Two sets
The different of A and B is the set of elements which belong to A but not to B and is
denoted by A/B
In example 1.2 above A/B = {x: x €A, x € B} Ω
A/B = {1}
B/A = {2}
The union of two sets A B can be divided into three disjoint sets
4
A/B, A ∩B, B/A.
That is
A ∩ B = (A/B) ∩ (A ∩B) ∩ (B/A)
The number of element in a set will be denoted by nA. Thus,
n(A ∩ B) = n (A/B) + n(A∩B) ∩ (B/A)
Since
(A/B), (A ∩ B), (B/A) are disjoint sets
nA = n(A/B) + n(A∩B)
nB = n(B/A) + n(A∩B)
n(A/B_ = nA –n (A ∩B)
n (B/A) = nB-n(A∩B)
hence,
n(A ∩B) = nA-n (A ∩B) + n(A ∩B) + nB-n (A ∩B) = nA + nB – n(A ∩ B)
note: A/B = A ∩ Bc, B/A = B ∩ Ac
for any set A,
A ∩ Ac = Φ, A ∩ Ac = Ω.
For any two sets A and B, we have the following decomposition:
B = B ∩ Ω = B ∩ ( A ∩ Ac) = (B ∩ A) ∩ ( B ∩ Ac),
Since A ∩ B and Ac ∩ B are disjoints, we have
nB = n(A ∩ B) +n(Ac ∩ B)
de Morgan’s Law
(i) For any two sets A and B
(A ∩ B)c = Ac ∩ Bc
5
Proof:
Ac = ( A ∩ B)c ∩ (B/A),
Bc = ( A ∩ B)c ∩ (A/B).
Thus,
Ac ∩ Bc = (A ∩ B)c
Thus
Series
A sequence is a set of numbers occurring in order, and there is a simple rule by which the
terms are obtained. For example, 1, x, x2… is a sequence. If the terms of a sequence are
considered as a sum, for instance, 1 + x + x2 + …. The expression is called a series. A
series with a finite number of terms is called a finite series otherwise it is called a finite
series otherwise it is called an infinite series. The summation is shown by the symbol.
When the sum is taken from the first term (r = 1) to thee nth term (r = n).
Example
Xr-1 = 1 ' () x =1
1–x
N, x =1
The commonratio is x. if – 1 < x < 1, xn -> 0 as n- > =, then the series converges to
1–x
Example
Solution
(1 – p)r-1 = 1 = 1
1 –(1 –p) p
And
(1 – p)r-1 = 1 – (1 – p)n = 1 –(1 –p)n
1 – (1 – p) P
Exponential series
The function y = ex is called an exponential function. This function is one of the special
functions of analysis and can be defined as the sum of an infinite series. It is defined by
Ex =1 + x + x2 +…+ xn+ … = xr
2! n! r!
The notation r! is called factorial r and is defined as
R! = r(r – 1) (r-2) ….. 2.1
For example
7
5! = 5 x 4 x 3 x 2 x 1
Where 0! = 1 and 1! = 1. R! defined above has no meaning unless r itself is a positive
integer.
The factorial of any negative integer is infinite.
Gamma Function
It is easily shown by direct integration that when m is an integer
M! =
Provided that m > - 1. It can be proved that (2) will also have a meaning for fraction m.
for instance, it is known that
Thus,
Generalization of the Bionomial theorem is the multinomial theorem. If n is a positive
integer
(a1+a2 + … ak)n = ∑…….∑ n! a1n1a2n2…aknk
N1 ! n2!....nk!
Sum is over all n1, n2….. nk where n1 + n2 + … + nk= n.
8
Product Notation
The product of the terms of a sequence x1, x2…., xn can be written as x1x2x3….xn. This
product is shown by the symbol.
Where the product is taken from x1, (the first term) to the nth term, xn. for example,
N1! N2!...nk! and a12…ak
Can be written as
Example
(a1 + a2 + a3)2 = ∑∑ 2!
3
Where n=2, k = 3. Possible values of n1 are
n1 =2, n2 =0, n3=0
n1 = 0, n2 = 2, n3 = 0
n1 = 0, n2 = 0, n3 =2
n1= 1, n2=1, n3 = 0
n1 =1, n2 = =, n3 =1
n1 = 0,n2 =1, n3 =
hence, we have
(aa + a2 + a3)2 = a12 + a22 + a23 + 2a1a2+ 2a1a3+2a2a3
Exponential Functions
These are function in which the variable occur in the index, for example ex32x are called
exponential functions, Let y = e h(x) then
For example, if y = e3x2 then
dy =6xe3x2
dx
Derivative of loge h(x)
If y = loge 2 x3 then
9
dy = 6x. 1 = 3
dx 2x3 x2
10
UNIT TWO
MATHEMATICS OF COUNTING
Probability had its origin in games of chance such as disc and card games. The number of
definitions you can find for it is limited only by the number of books you may wish to
consult. Probability can be defined as a measure put on occurrence of a random
phenomenon. probability theory is developed as study of the outcomes of trail of an
experiment.
Definition
An experiment is a phenomenon to be observed according to a clearly defined procedure.
Probabilities are numbers between 0 and 1, inclusive that reflect the chances of a
particular physical events occurring. If a die is tossed once, the possible outcomes are
1,2,3,4,5,6, let Ω = {1,2,3,4,5,6}. Then Ω is a set consisting of all possible outcome of
tossing a die once. This set is given a name, it is called a sample space.
Example 1.1
Write down the staple space for each of the following experiments
(i) Toss a coin 3 times and observe the total number of heads
(ii) A box contains 6 items of which 2 are defectives. One item is chosen one after the
other without replacement until the last defective items is chosen. We observe
the total number of items removed from the box.
Solution
(i) Possible outcomes are: number of heads is 0 when we have TTT, number of heads
is 1 when we have HTT or THT or TTH, number of heads is 2 when we have
HHT, THH, HTH and number of heads are 0,1,2,3. Hence the sample space Ω
= {0,1,2,3)
(ii) Total number of items removed is 2 if we have “the first is defective (D) and the
second is defective D (Note that the total number of items removed can not be
0 or 1) denoted by DD. The total number of items removed is 3 if we have
11
GDD or DGD (where G denoted good item), and so on. Thus, the sample space
Ω = {2,3,4,5,6}
Definition 1.2 Event
Any subset A of a sample space Ω is called an event, where A < Ω.
Example 1.2
The following are examples of events
(i) An odd number occurs when a die is rolled once
Ω = {1,2,3,4,5,6}, A = {1,3,5}
(ii) Toss a coin 3 times, the total number of heads observed is even
Ω = {0,1,2,3} A = {0,2}
Example 1.4
A coin is rolled thrice. The possible outcomes assumed to be equally likely are
Ω= {HHH, HHT, HTH, HTT, TTT, TTH, THT, THH}
Let A be the event that 2 heads occur. Then
A = {HHT, HTH, THH}
n Ω = 8, nA = 3
hence
P(A) = 3
8
This classifiableapproach when applicable (the possible outcomes are equally likely) has
the advantage of being exact. Thus to compute a probability by using the above definition
one must be able to count.
(i) n Ω the total number of possible outcomes in the sample space and
(ii) nA, the number of ways in which event A can occur
equally like outcomes are also called equally probable outcomes. If an event cannot occur
its probability is 0, if it must occur its probability is .
the computation of nA and n Ω is easy if Ω has only a few possible outcome, as the
number of possible outcomes becomes large, this method of counting all possible
outcomes are cumbersome and time consuming.
Alternative methods of counting must therefore be developed. For example if one asks
for the probability of getting sum of numbers showing to be 30 when 7 dice are rolled
13
one must determine how many different ways are possible to get the sum to be 30. Such
possible ways include
(6,6,6,6,2,1,3) (2,5,2,6,6,3), (4,4,4,4,4,4,6)
In this chapter weintroduce nontechnical discussion of techniques of mathematics of
counting frequently needed in problems of finding nA and n Ω
Exercise 1.1
1. A die is rolled once. What are the probabilities of getting
(i) An even number
(ii) A prime number
(iii) An off prime number
(iv) An odd number
(v) An even number
Fundamental Principle of Counting
First Law of Counting
If an event A1 can occur in n1 ways and thereafter an event A2 can occur in n2 ways,
“both A1 and A2 can occur in this order in n1n2 ways.
Example 1.5
Roll a die first and then a con
A1 can occur in 6 ways (1,2,3,4,5,or 6) and A2 can occur in 2 ways (H or T).
Thus, by the above law, there are
6 x 2 =12
Possible ways for the outcomes
The outcomes are
(1,H) (1,T) (5, H) (5,T)
(2,H) (2,T) (6,H) (6, T)
(3, H) (3, T) (4, H) (4,T)
In general, if an event A1 can occur in n1 in different ways and if following this an event
A2 can occur in n2 different ways, and if following this second event, an event A3 can
14
occur in n3 different ways and so forth, then the events A1 and A2 and A3…. And Ak can
occur in this order in n1n2….nk ways.
Example 1.6
If a die is rolled 10 times. Let A1denote the outcome of the roll, i= 1,2 ,…. 10. A1 can
occur in 6 ways.
Thus thenumber of ways A1, A2,…. And A10 can occur is 610possible outcome.
That is
N Ω=610
Second Law of Counting
If an eventA1 can occur in n1 different ways and an event A2 can occur in n2 ways then
either A1 orA2 can occur in n1 +n2 different ways.
Example 1.7
Let us toss a die or a coin once. let a1 ne the event “the die shows an even number and a2
be the event “the coin lands heads”. a1 can occur in 3 ways (2, 4or6) and a2 can occur
only ocne.
\the number of ways in which an even number or a head be obtain is
3+ 1 =4
A1 can occur in 3 ways A2 can occur in 1 way
Therefore A1 or A2 can occur in 4 ways
In general, if events A1, A2….,Ak can occur in n1….nk different ways, then either or A1
+ n2 +… + nk different ways
Example 1.8
Four people enter a restaurant for lunch in which there are six chars. In how many ways
can they be seated.
Let A1, A2,A3,A4 denote the events “choice of chair by the four people”, the suffix
denoting order of seating. The first person to sit down has six choices. He can decide tosit
on any of the six vacant chairs. Therefore, there are 6 different ways A1 can occur, after
15
the first person has seated, the second person can sit on any of the remaining 5 chairs.
Thereafter the third person can sit on the remaining 4 chairs.
Thus, using the notation of law of counting
N1 =6, n2 = 5, n3 = 4, n4 = 3
6 x 5 x 4 x 3 =360 ways
They can be seated .
Example 1.9
How many 4 digits numbers can be formed from the digits 0,1,2,3,4,5 if the first digit
must not be 0 and repetition of digits are not allowed.
Let A1, A2, A3, A4, denote mthe events: select the first second, third and fourth digits
[Link] the first digits can not be 0, the first digit can e either 1,2,3,4, or 5
therefore A1 can occur in 5 ways, having chosen the first digit the second digit can be
selected from the remaining 5 digits. there the first and second digits have been chosen,
there remains 4 digits. The third digit can be chosen from the remaining 4 digits and the
fourth can be chosen from the remaining 3 digits. Therefore,A1, A2, A3 and A4can occur
in 5,5,4,3 ways respectively
Thus there
5 x5x4x3 = 300 numbers
Example 1.10
Ten candidates are eligible to fill 4 vacant position. How many ways are there of filling
them?
Let A1,A2, A3,A4 be the events denoting candidates that fill positions 1,2,3,4 respectively.
Candidate filling position 1 can be any of the tend candidates, therefore A1 can occur in
10 ways. Following this, the next position 2 can be filled by any of the remaining 9
candidates since position 1 had been filled by one candidate. Therefore A2 can occur in
ways. Similarly, A3 and A4 can occur in 8 and 7 ways respectively. Thus there are
10x9x8x7 = 5040 ways
Exercise 1.2
1. A student is to answer all the five questions in an examination. It is believed that
the sequence in which the questions are answered may have a considerable effect
16
on the performance of the student. In how many different order can the question
be answered
2. If a woman has 10 blouses and 6 skirts, in how many ways can she choose a dress
assuming any combination of blouse and skirt matches
3. In a study of plants, five characteristics are to be examined. If there are six
recognizable differences in each of four characteristics and eight, recognizable
difference in the remaining characteristics. How many plants can be distinguished
by these five characteristics?
4. A bus starts with 6 people and stops at 10 different [Link] many different ways
can the 6 people depart if
(i) Any passenger can depart at any bud stop
(ii) No two passengers can leave at the same bus stop
5. Show that the number of ways of choosing r objects from n objects with
replacement is given by n”
Permutation and Combination
A permutation is an arrangement of objects in a definite order (called ordered sample). A
combination is a selection of objects without regard to order (unordered sample)
A group of objects, with regard to permutation and combination has three characteristics:
1. The way the objects in a group are arranged
2. The kind of objects in the group
3. The number of objects of each kind in the group
1.2 permutations:
Suppose that a set contain n objects. We are often interested in arranging the objects in a
definite order.
Two groups containing n object are said to form different permutations if they differ in
arrangements
Consider groups of letters a, b, c, d.
(a, b, c, d): (b, c, d, a): (c, a, d, b)
17
They are all different permutations because the arrangement of the letters is different in
each group.
Example 1.11
How many permutations of four letters can be formed from the letter a, b, c, d.
To answer this question we reason as follows:
Since are permuting 4 letters, four events are involved. Let A1 be the events denoting the
letter to occupy the ith position. A1 can occur in 4 ways, that is the first letter can be
either a, b, c, or d. after this event has taken place, A2 can occur in three ways (that is,
after having chosen the first letter, the letter occupying the second position can be chosen
from the remaining three letters. After the second event, there remains only two letters
from which one is to be chosen tooccupy the third position. So A3is 2 and it remains only
one letter.
A4 can occur only in one way. Thus, there are 4 x 3 x 2 x 1 = 24 permutation of the four
letters. Permutation of four objects from four objects is called permutation 4 and is
denoted by 4P4
Example 1.12
If n balls are distributed at random into r boxes, in how many ways can this be done if
(i) Each of the balls can go into any of the r boxes
(ii) No box has more than one ball.
Solution
(i) Let A1, A2….An where the object occupying the ith position is the outcome of A1 I
= 1,2,…..n. A1 can occur in n1 way, A2 can occur in n-1 ways and so on.
18
(ii) The first ball can go into any of the r boxes, the second any one of the remaining
(r-1) boxes etc, so in all there are r (r-1) (r-2)… (r-n+1) different ways (n < r).
19
Example 1.13
How many permutation of three letters can be formed from the letters a, b, c, d, e. abc,
bae, cba, cdb,….. are few of the required permutation. Since the permutation consist of 3
letters, 3 events are involved, A1, A2 A3. The first letter can be any of the 5 letters,
therefore A1 can occur in 5 ways. Similarly A3 can occur in 3 ways. Thus, there are
5 x 4 x 3 =60
Permutation of three letters from 5 letters. This is denoted by 5P3 (5 permutation 3).
Definition 1.4
The number of permutation of r (r < n) objects from n district objects is called nn
permutation r and is denoted by npr
It can easily be shown that
n
Pr =n(n-1)(n-2)…(n-r+1)
to see this we can argue as follows. Since we are permuting r objects, there are r events
A1, A2…..A3 involved.
The first position can be occupied by any of the n objects, following this the second
position can be occupied by any of the remaining n-1 objects and so on. Therefore A1,
A2,…Ar can occur in n, n-1, n-2, …n-(r-1) ways respectively. thus the number
ofpermutation is
n(n-1)(n-20…(n-r+1)
But
n!= n(n-1) (n-2)…(n-r+1) (n-r)(n-r-1)...1 (3.2.1)
(n-r)! = (n-r)(n-r-1)…2.1
n! = n(n-1)(n-2)…(n-r+1).
(n-r)!
20
Therefore
n
Pr=n(n-1) (n-2)…(n-r+1)
n
Pr= n!
(n-r)!
In example 1.13 above, n = 5, r = 3
5
P3= 5! = 5! = 5 x 4 x 3 = 60
(5-3) 2
Example 1.14
A group of students consist of 5 men and 3 women. The students are ranked according to
their performance in a quiz competition. Assuming no two students obtain the same score
(i) How many different ranking are possible?
(ii) If the men are ranked just among themselves and the women among themselves,
how many different rankings are possible?
Solution
(i) A possible ranking corresponds to a permutation of the students. The number of
possible permutation of the 8 students gives the number of different ranking
possible. Thus the answer is
8
P8 = 8! = 8! = 40320.
(8-8)!
(ii) There are 5P5 = 5! = 120 possible rankings of the men and 3P3 =3! = 6 possible
rankings of the women. It follows from the fundamental principle of counting
that there are
5
P5 x 3P3 = 5! X 3! = 72- possible rankings
Example 1.15
Four digits numbers are to be formed using any of the digits 1, 2, 3, 4, 5, 6, (No repetition
of digit is allowed).
(i) How many four digit number can be formed
(ii) How many 4 digit numbers greater than 3000 can be formed?
(iii) How many 4-digit even numbers can be formed?
21
(iv) How many of these even 4-digit numbers are greater than 3000?
Solution
(i) This is permutation of 4 digits from 6 digits. Therefore the different permutation is
6
P4 = 6! =360.
2!
Alternatively, we can reason as follow:
The first digit to be selected can be any of the six given digits, so n1 = 6. The second
digits to be select can be any of the remaining 5 digits (since no reparation is allowed) so
n2 = 5. Similarly, n3 = 4 and n4 = 3. Thus, the answer is
N1 x n2 x n3 x n4 = 6 x 5 x 4 x 3 = 360
(ii) Since the number must be greater than 3000, the first digit must be chosen room 3,
4, 5 or 6 so n1 = 4. The second digit can be any of the remaining 5 digits, so n2
=5 similarly, n3 = 4, n4 = 3. Thus, there are
4 x 5 x 4 x 3= 240
4 digit numbers greater than 3000 that can be formed using the digits 1, 2, 3, 4,5, 6,
(iii) A number can be defined to be an even number of the last digit of the number
is even. Going by this definition, we see that the required number is even if the
last digit is 2,4, or 6. Since there is a restriction on the last digit. We have to
select the last digit first. The last digit can be selected from 2,4 or 6, so we
have 3 choices. Having chosen the last digit, we can now select the first,
second and third digit. After chosen the last digits there remains 5 digits, so n1
= 5 similarly, n2 = 4,n3 =3. Thus, the answer is 5x 4 x 3x 3 = 180
(iv) The last digit must be 2, 4 or 6 and the first digit must be 3, 4, 5or 6. If the last
digits is 2, then the number of ways of selecting the first digit is 4. However, if
the last digit is selected is 2 or 6 the number of ways of selecting the first digit
is 3 (that is 3, 5 or 6, 3, 4 or 5)
Case 1: n4 = 1 n1 = 4 n2 =4, n3=3
The number of 4 digits numbers that can be formed in this case is
4 x 4 x 3 x 1 = 48
Case II: n4 =2, n1 = 3n2 = 4, n3 = 3
22
The number of 4 digit numbers that can be formed in this case is
3 x 4 x 3 x 2 = 72
Hence, the total numbers of even 4 digit numbers greater than 3000 that can be formed is
48 +72= 120.
Example 1.16
The letters A,B,C, D and E are placed at random to form a five letters word (without
repetition). How many ways can a word be formed such that
(i) D directly follows A.
(ii) A and D follow each other
(iii) A, D and E follow each other
Solution
(i) For D to directly follow A, we must always have AD appearing in the word so
formed AD can be regarded as a letter so that we now have the letters B,C E
and AD. The number of ways of rearranging tehse new four letters is 4! Thus
there are 4! = 24 ways of forming a word such that D directly follows A
(ii) A and D follow each other; either we have AD or Da each having 4! Ways. Thus
there are 2x 4! Ways of forming a word such that A and follow each other
(iii) ADE can be regarded as a letter so that we now have the letters B.C ADE.
There are 3! Permutation of the new 3 letters. ADE can also be permuted in 3!
Ways, that is
ADE, AED, DAE, DEA, EAD, EDA.
Thus there are
3! X 3! = 36 ways
Permutation of Indistinguishable Objects
Consider n objects where n1 are of type 1, of type 2,…. Nk of type k. ,in shown many
ways can the n object be arranged.
For example in how many ways can the letters of the word book be arranged. here n= 4, k
= 3, n1= 1, n2= 2, n3 = 1. First give the two O’s suffixes bo1o2k. then treating the O’s as
different, the 4 letters may be arranged in 4! Ways. In every distinct arrangement, the
23
2O’s may be rearranged amongst themselves in 2! Ways without altering the permutation
for instance o1bo2k are the same when the suffixes are removed.
Therefore, the number of permutations of the letters of the word book is
4! = 6
2!
Book, ookb, oobk, obok, okob, koob.
In general, the number of ways in which n objects where n1 are of type 1, n2 of type 2…,
nk of type k can be arranged is given by
n!: n + n2 +…+ nk
n1!n2!...nk!
Example 1.17
How many distinct permutations are there of the letters of the word Television?
The ten letters to be permuted consist of 2e’s, 2i’s,IT, I, v, s, o, n. thus the number of
distinct permutation is
10! = 10!
2! 2! 1! 1! 1! 1! 1! 1! 2! 2!
1.3 Combinations
Definition 1.5
Two groups are said to form different “combination” if they differ in the number of
any kind of object in the groups. Consider a group of 4 letters a, b, c, d. the
combination abcd, bcda, cadb are identical combinations each of them contains the
same number of a, b, c, d: one b, one c and one d.
The combinations of the 4 letters taken 3 at a time are:
Abc, acd, abd, cbd
Therefore, there are 4 distinct combination for three letters from the four letters. Each
of tehse combinations has 3! = 6 permutation. For instance.
abc = abc, acb cab cba bac bca
acd = acd adc cad cda dac dca
abd = abd adb bad bda dab dba
24
cdb = cbd cdb bcd bdc dbc dcb
The number of distinct permutation is 4P3=24
The number of distinct combinations is 4.
Therefore.
Number of combinations = number of permutations
3!
The number of distinct combination of 4 objects takn 3 at a time is denoted by 4C3.
Thus,
4
C3 =4P3/3!
In general,
n
Cr =nPr/r!= n!
(n-r)!r
Theorem 1.1
The number of distinct combinations of n objects taken r at a time (that is the number
of ways of choosing r objects out of n, disregarding order and without replacement) is
given by
Proof:
There are nPr = n!
(n-r)! permutations of n objects taken r at a time. If we disregard
order among the r objects, there are r! permutation that will give the same
combination. Therefore the number of combinations is the number of permutation
divided by r! thus.
n
Cr = nPr = n!
r! (n-r)!r!
Example 1.18
A club consist of 15 members. In how many ways can a committee of 3 to be chosen?
Solution
This can be done in 15C3 ways
15
C3 = 15! = 15 x 7 x 13 = 455 ways
25
3! 12! 3
Example 1.19
A club consist of 10 men and 5 women, in how many ways can a committee of 6
consisting of 4 men and 2 women be chosen. The 4 men can be chosen from the 10
men in10C4 ways, the 2 women can be chosen from the 5 women in 5C2 ways. Hence
the committee can be chosen in (by the fundamental principle f counting)
10
C4 x 5C2= 10! X 15! = 2,100 ways
6!4! 2!3!
Exmaple1.20
Suppose we have a box containing n balls of which r are black and the remaining
white. A random sample of size k is selected without replacement. In how many ways
can the sample be selected such that it contains x black balls
The x black balls can be selected in rCxways and k-x white balls n be selected from n-
r white balls. Thus, there are
r
Cx Xn-rCk-r
ways of selecting a sample such that it contains x black balls and k-x white balls
Example 1.21
A committee of 4 men and 2 women is selected from 10 men and 5women. If two of
the men are feuding and will not serve on the committee together, in how many ways
can the committee be selected
Solution
The number of ways of selecting 4 men and 2 women is
10
C4 x 5C2 =2100
The number of ways of selecting the committee such that the two men are in the
committee is
8
C2 x 5C2 =280
26
Hence, the number of different committees that can be formed such that the two men
are not in the committee together is
2100 – 280 = 1820
Another method is to consider 3 cases
Case I: The two men say, A and B are not in the committee
8
C4 x 5C2= 700
Case II: A is the committee but not B:
8
C4 x 5C2 = 560
Case III: B is the committee but not A
8
C4 x 5C2 = 560
Thus, the total number of ways the committee can be selected is
700+ 560+ 560 = 1,820
Example 1.22
How man subsets can be formed, containing at least one member from a set of n
element
Solution
There are nCk subsets of size k that can be formed, the total number of subsets
containing is least one member is
Example 1.23
A student is to answer 5 out of 8 question in an examination. How many if he must
answer at least 2 of the first for 4 questions?
Solution
(i) By the combination law the answer of 8C5 = 56
(ii) Possible choices are (2,3) (3,2) (4,1) where (2,3) means answer 2 questions from
the first 4 questions and 3 from the remaining 4 questions. The number of ways
of doing this is 4C2 x 4C3.
Thus the answer is
27
(4C2 x4C3) + ( 4C3 x 4C2) + (4C4 x 4C1)= 52
1.4 Partitioning
Dividing a population or a sample of n objects into k ordered parts of which the first
contain r object, the second r2 objects and so on is called ordered partition. If the
division is into k unordered parts, then the partition is said to be unordered. For
example suppose a class contains 15 students and we want to divide the class into 3
tutorial groups of 5 each. Three lecturers are available and each is to take each group.
In otherwise, we want to divide the 15 students into 3 ordered groups (A,B,C) this is
an ordered partition since there are
3! = 6 ways
The lecturers can be assigned to take any partition, for instance groups (A, B, C) can
be taken by
(L1, L2, L3) or (L2, L1, L3) or (L1, L3, L2)
Or (L2, L3, L1) or (L3, L2, L1) or (L3, L1, L2)
Where L1 means lecturer I ad (L1, L2, L3) means L1 takes groups A, L2 takes group B,
L3 takes group C and so on
15
There are C5 ways of selecting those students those students to be in group A,
following this, there are 10 students left and so there are 20C5 selecting those to be in
the second group, therefore by the fundamental principle of counting, there are
15
C5 x 10C5x 5C5= 15! X 10! X 5!= 15!
10! 5! 5! 5! 5! 5! 5! 5!
Ordered partitions. The number of ways in which n object can be divided into ordered
parts of which the fist contains r1 objects the second r2objects and so on is
n
Cr1. n-r1Cr2.n-r1-r2Cr3…n-r1-r2-…-rk-1Crk
28
Example 1.24
In how many ways can three committees of five, three and two persons be formed
from 10 persons.
We seek the number of ordered partitions of the 10persons
This is given by
10! =2520
5! 3! 2!
Example 1.25
In how many ways can 9 toys be divided among three children if each gets 3 toys
If the toys are numbered 1 through 9, the partition {(1, 2, 3), (4, 5, 6), (7, 8, 9}) means
child A gets toys 1, 2, 3 child B gets toys (4, 5,6) while child C gets toys 8, 8, 9. We
distinguish between {(1, 2, 3), (4, 5,6) (7, 8, 9)} and {(4, 5, 6), (1, 2, 3), (7,8, 9) so
these are ordered partitions. Thus there are
9! = 1680 ways
3! 3! 3!
When r1 = rj for I = j, we can distinguish between ordered and unordered partition for
example partition a set consisting of 8 objects numbered 1 to 8 into 3 parts of which
the first contains 2 objects, the second 4 objects and the third 2 objects. A partition is
{(1, 2), (3, 4, 5,6), (7, 8)}.
In an ordered partition we distinguish between the partition
{(1,2), (3,4, 5, 6), (7, 8)} and {(7, 8), (3, 4, 5, 6), (1, 2)}
But they are the same for unordered partition. Therefore, the number of unordered
partition is
8! . 1 = 210
2! 4! 2! 2!
Example 1.26
29
In how many ways can a family of 9 divide itself into 3 groups so that each group
contains 3 persons?
Solution
We are seeking for unordered partitions r1 = 3, r2=3 r3=3
The number of unordered partitions is
9! 1 = 280
3! 3!3! 3!
Since the three parts contain the same number of objects
The same relationship that exists between permutation and combination exist between
ordered and unordered partition of a set.
Example 1.27
In how many ways can a family of 10 be divided into three groups, one containing 4
and the others 3?
Solution
R1=4,r2 =, r3=3
The number of ordered partitions is
10! =4,200
4! 3! 3!
And the number of unordered partition is
10! 1 = 2,100
4! 3! 3! 2!
Since only two of the three parts contain the same number of objects
30
Now, suppose the balls are non-district (indistinguis-shable) we can only talk about
number of balls in the ith cells
Let X1, x2…, xn denote the number of balls in the ith cell, then
x1 + x2 +…+xn = r
The number of distinct distribution in which no cell remains empty is
r-1
Cn-1
to see this let us assume that the r non-distinct objects are lined by and bars used to
divide them into groups. The r balls r-1 space of which n-1 are to be occupied by bars.
For example if r =9 and n = 5, we have
0∩0∩0∩0∩0∩0∩0∩0∩0
Thus 0/00/0/000/00 corresponds to
X1 =1,x2 = 2, x3 = 1, x4 = 3, x5 =2.
And there are 8C4 possible distribution of the bars. another possible distribution is
000/0/00/00
If x1> 0, that is cells can remain empty, the number of non-negative solutions of
X1 + X2 +…+ xn=r
Is the same as the number of positive solutions of
Y1 + y2 +…+ yn =r + n
Where y1 = x1 +1. Thus, there are
r + n-1
Cn-1 (1.2)
Distinct solution satisfying x1 + x2 +… + xn = r
The number of distinct distribution is the number of ways n – 1 spaces can be selected
out of the n + r-1 spaces
Application to Runs
Definition
A run is any ordered sequence of elements of two kinds
For example, by a run of wins we mean a consecutive sequence of wins. The
sequence WWWLWWLLWLWW gives 4 runs of wins. The first run is length 3, the
second run of length 2, the third of length 1 and the fourth of length 2.
31
Suppose now that we have two letters (W and L) n non-distinct letters (L) and m non-
(n+m)
district letter (W). the total number of distinct orderings of W and L is Cn-r r runs
of W is equivalent to arranging the letters W into r cells none of which is empty. If
there are r runs on W,
The number of L run is necessarily r+ 1, r-1 or. Thus from (1.1) we have (m-1)Cr-
1distinct was of having r runs of W. hence there are.
(m-1)
Cr-1(n-1)Cr ways of having r runs of W and (r + 1) runs of 1…
A sequence representing r runs of W is
33
14. A disciplinary committee of four is to be chosen from six men and five women.
One particular man and one particular woman refuses to serve if the other person
is on the committee. How many committees may be formed
15. Eleven people are to travel in two cars- salon and station wagon. The saloon has 4
sets and the station going 7 seats. In how many ways can party be split up?
16. In how many ways can a committee of 6 composing of 3 full professors, 2
associate professors and 1 senior lecturer be selected from 5 full professor, 10
associate professors and 20 seniors lecturers
17. A box contain 12 balls labeled 1, 2, 3… 12, suppose a random sample of size 4 is
selected. In how many ways can the sample be selected if balls labeled 2, 3, are
among the four selected.
18. Suppose a random sample of size r is drawn from a population of n objects. In
how many ways can this be done if k given object must be included in the sample
and (a) sampling is without replacement (b) Sampling is with replacement
19. I bought 2 tickets to a lottery for which n tickets were sold and 4 prizes to be given
in how many ways can the tickets be drawn such that 1 win at least a prize?
20. A committee of 8 is to be formed from 10 couples (10men and 10 women). In how
many ways can the committee be formed if no husband serves on it with his wife.
21. Lines are drawn to pass through six points. In how many ways can this be done if
each line passes a through only two points?
22. Interchanges may occur between any two of the n chromosomes of a cell
(a) In how many ways can exactly one interchange occur?
(b) In how many ways can exactly k interchanges occur?
(c) If n = 5, in how many ways can at most three inter changes occur?
23. Prove that (n1-n2)Ck = n1Crn2Ck-r where k < n1, n2
HINT: Select k objects from n1 + n2 obejcts of n1 of type 1 and n2 of typeII
24. A company is considering building additional warehouse at new locations. There
are ten satisfactory location and the company must decide how many and which
ones to select. How many choices are there?
34
25. Are there more samples obtainable in five draws from 10 objects with replacement
than 12 objects without replacement?
Each of fifty items is tested and found to be defective or non-defective. How many
possible outcomes are there?
MODULE TWO
UNIT ONE
Example 2.1
Suppose a fair die is rolled once. There are six possible outcomes. The sample space is
Ω = {1, 2, 3, 4, 5, 6}.
35
A2 = “an odd number occurs
The six possible outcomes in Ω are equally likely. The probability that A occurs is
denoted by P(A) and the probability that A does not occur by P(AC). if we let nA be the
number of outcomes that have attribute A, then
A1 = {2, 4, 6}
A2 = {1, 3, 5}
A3 = {2, 3, 5}
Thus,
nA1 3 3
P( A1 ) = = 3 / 6 = 1 / 2, p ( A2 ) = , P ( A3 ) =
nΩ 6 6
Example 2.2
Suppose that a box contains 10 items of which are defective. Two items are selected at
random without replacement. Find the probabilities that:
(i) both items are non-defective (ii) only one item is defective
(iii) both items are defective: (iv) at least one item is defective
10
C 2 = 45 ways.
36
So there are 45 elements in the sample space.
(i) The number of ways of selecting 2 items from the non-defective items is 6C2 = 15.
That is A1 can occur in 15 ways. Thus.
/ / 173
*+ 01- 34
,
/
-. + 02 54
P(A1) =
,
(ii) The number of ways of selecting 1 item from the 4 defective items and 1 item
from the 6 non-defective items is 4C1 × 6C1 = 24 so, A2 can occur in 2 ways.
Thus,
/ /
8+ 9 * + :5 ;
- -
-. + 54 34
,
P(A2) =
Similarly
/ /
8+ < :
,
-. + 54 34
,
(iii) P(A3) =
= / 273.
; :
34 34
P(A4) = P(1 defective item) + P(2 defective items) = P(A2) + P(A3) =
Example 2.3
Suppose a fair die is rolled twice. Find the probability that the sum of the numbers on the
two faces is (i) even, (ii) less than 5.
(i) Let A be the event “the sum of the two faces is even”. Possible outcome are:
37
/ 1⁄2
3;
><
P(A) =
(ii) Let B be the event “the sum is less than 5”. B occurs if the sum is 2, 3, or 4. The
sum is 2 if the outcome is (1,1) the sum is 3 if the outcome is (1,2) or (2,1) and the
sum is four if the outcome is (3,1), (1,3) or (2,2). Therefore B can occur in
1 + 2 + 3 = 6 ways
Thus
/ 176.
<
><
P(B) =
Example 2.4
3 balls are drawn at random with replacement from a box containing 8 red and 3 white
balls. Find the probability that (i) all 3 are red; (ii) 1 is red and 2 are white.
The sample space consists of 113 possible outcomes. Let A be “the evet all 3 are red” and
B the event “1 is red and 2 white”
/
01 ;A
02 33A
P(A) =
(The first red can be chosen in 8 ways the second in 8 ways and the third Red in 8 ways).
0B
02
P(B) = , nB = 8,3.3 + 3.8.3 + 3.3.8
(8 × 3 × 3) = number of ways of picking the first to be red, second white and third white
(RWW), 3 × 8 × 3 is for WRW and 3 × 3 × 8 is for WWR). Thus
.
>.;.>.>
33A
P(B) =
That is, two or more events are mutually exclusive if no two of them have points in
common. In example 2.1
A1 = {2, 4, 6}
A2 = {1, 3, 5}
A1 A2 = ,
Definition 2.1
If A1 and A2 are any two events, A1 A2 is the event that occurs if and only if both A1
and A2 occur. In general, if A1, A2, …, An occur.
A1 A3 = {2}
Thus
Thus
P(A1 A2) =
Complementary Event
39
The event “A occurs” and “A does not occur” are mutually exclusive. The event “A does
nt occur” is called the complement of A and is denoted by AC.
Theorem 2.1
If AC is the complement of an event A, then P(Ac) = 1 – P(A). This theorem states that the
probability that an event will not occur is equal to 1 minus the probability that it will
occur.
In example 2.2 (iv); The event “no defective item” is the same as the event both items are
non-defective”.
P(A4) = 1 – P(
A4c = A1 since the event “no defective item” is the same as the event” both items are non-
defective.
Definition 2.2
If A1 and A2 are any two events, A1 A2 is the event that occurs if t least one of A1 and
A2 occurs.
In general, if A1, A2,…, An are any events, A1 A2 is the event that occurs if at least one
of A1 i = 1, 2,…, n, occurs.
In example 2.1, the event A1 A3 is the event that an even or a prime number occurs.
A1 A3 = {2, 3, 4, 5, 6}.
40
Thus
P(A1 A2 ) = =
A1 A2 = {1, 2, 3, 4, 5, 6} = Ω,
thus
Theorem 2.2
Corollary
A1 A2 = ϕ
hence
41
P(A1 A2 ) = 0
(iii) If A and B are any two events defined on the sample space then
Theorem 2.2 can be generalized for n > 2 events. It can easily be shown that for any 3
events A1, A2, A3
42
(by theorem 2.2). thus
Note:
Definition 2.3
Two or more events are said to be exhaustive if their union equals the whole sample
space. In other words, A1, A2,…, An are exhaustive events if
P(A1 A2 … An) = 1
Note:
1. 0 ≤ P(A) ≤ 1
2. P(Ω) = 1.
Examples:
2.5 A box contains 6 balls numbered 1 to 6. A ball was drawn from the box at random.
Find the probability that the number on the ball drawn was either 1, 2, or 6.
Solution
43
Let A1, A2, A3 denote the events that the ball drawn was 1, 2 and 6 respectively.
A1 A2 A3 denote the event the number on the ball drawn was either 1, 2 or 6.
A1, A2 and A3 are mutually exclusive events. Thus
2.6 Four fair dice are tossed once, what is the probability that the sum of the numbers
on the four dice is 23?
Solution
Therefore, the probability that the sum of the numbers on the four dice is 23 is
4/64.
2.7 If P(A1) = 2/3, P(A1 A2) = ¼ and P(A1 A2) = 5/6, find P(A2).
Solution
2.8 Suppose a fair on is tossed three times, what is the probability that at least one
head occurs?
Solution
44
Let A1 be the event that the first toss lands heads, A2 the event that the second toss
lands, and A3 the event that the third toss lands heads.
A1 A2 A3 is the event that at least one head occurs.
A2 = A1 (A2/A1).
Since A1 and A2/A1 are mutually exclusive, we have
P(A2) = P(A1) + P(A2/A1)≥ P(A1)
Since P(A2/A1) ≥ 0.
2.9 If A1, A2,…, An are n events, then
=
From de Morgan’s law, we have
=
hence
45
the probability that an event A will occur “conditional on” the knowledge that another
event B has occurred.
Suppose a fair die is rolled and it is known that an even number appeared uppermost. Let
A be the event that the number was greater than 3 and B the event that the number that
appeared was even. The problem is to find the conditional probability that the event A
occurred given that the event B has occurred, P(A|B). since we know that the number was
even, the number must be either 2,4 or 6. Therefore, the conditional sample space
contains 3 elements. The event A occurs if the number showing is 4 or 6, thus
P(A|B) = 2/3.
We can therefore define conditional probability of A given Bas the number of ways AnB
can occur divided by number of elements in the conditional sample space. That is,
,
Where ≠ 0, where ΩB is the condtional sample space given that B has occurred.
Divide the numerator and denominator by nΩ
Examples
2.11 Suppose a box contains 4 red balls and 3 black balls. Compute the probability that
(i) the second ball drawn is red if the first ball drawn was red; without
replacement,
(ii) the second ball drawn is red if the first ball drawn was black
46
Solution
If the first ball drawn is red, there remains 6 balls, 3 red balls and 3 black balls.
The probability of the second ball being red is 3
6 = 12 . But if the first ball is black,
the box is left with 4 red and 2 black so the probability of the second ball being red
is then 4/6 = 2/3. Thus
This shows that the probability of the event “the second ball drawn is red” depends
on the colour of the first ball drawn.
2.12 Suppose two fair dice are rolled. If the sum of the numbers appearing is 6, what is
the probability that one of the number is 2?
Solution
Let A be the event “one of the numbers is and B the sum is 6. there are five ways
for the event B to occur: (3,3), (2,4), (4,2), (5,1) and (1,5) and there are two ways
for the event AnB to occur: (2,4) and (4,2).
Thus,
P( A ∩ B) = 2 36 , P( B) = 5 36
Hence,
P( A ∩ B)
P( AΙB ) = = 2 5.
P( B )
47
2.13 There are two children in a family. If there is at least a girl in this family, what is
the conditional probability that both are girls.
Solution
Ω = {BB , GB , BG , GG}
Let A be the event “both children are girls and B ”a least a girl in the family.
P ( AnB ) 1 4 1
P(A Ι B) = = = .
P( B ) 34 3
2.14 There are three children in a family. If there is at least one boy and at most two
boys in this family. What is the conditional probability that there are exactly two
boys in this family.
Let B be the event “at least one boy and at most 2 boys in the family” and let A be
the event “exactly two boys in the family”. Then
Therefore
P ( AnB ) 3 8 3 1
P(A Ι B) = = = = .
P( B ) 68 6 2
Exercises 2.1
(ii) If it is known that the difference the two numbers was 3, what is the
probability that the sum of the two numbers was 7?
2. Two unbiased dice are thrown once. What is the probability that
3. Suppose events A and B are sun that P(A) = 1/5, P(AnB) = 1/6.
4. If A and B are two events defined on the same probability space, show that: (i)
P(A) = P(A∩B) + P(A ∩ Bc) = P(B) + P(A ∩ Bc) – P(Ac ∩ B).
(i) both coins show a tail given that the first shows a head;
(ii) both are heads given that at least one of them is a head.
7. A red die and a green die are rolled once. Find the conditional probability that:
(i) the number on red die is odd, given that the sum of the two numbers
showing is 9;
(ii) the sum of the two numbers is 9 given that one of the numbers is odd and
the other even?
49
2.3 Bayes Theorem
suppose a box contains r red balls and b black balls. Two balls are drawn at random
1
without replacement. Assume that the probability of drawing any particular ball is r +b
.
Let A1 be the event “the first ball drawn is red and let A2 be the event” the second ball
drawn is red. Then
r r −1
P( A1 ) = , P( A2 Ι A1 ) =
r +b r + b −1
(
P A2 Ι A1 =
c
) r
r + b −1
A2 = A1 ∩ A2 or A1c ∩ A2.
Therefore,
P(A2) = P(A2 ∩ A1) + P(A2 ∩ A1c) = P(A1) P(A1) P(A2 Ι A1) + P(A1c) P(A2 Ι A1c).
Theorem 2.3
k k
= ∑ P( B ∩ A ) = ∑ P( A ) P( BΙA ).
i =1
1
i =1
1 1
Example 2.15
Suppose a box contains 3 red balls, 2 black balls and 5 green balls. Two balls are drawn
at random without replacement. Find the probability that the second ball drawn is red. Let
A1 be the event “the first ball drawn is red A2 the event the first ball drawn is black and
50
A3, the event “the first ball drawn is green. Let B be the event “the second ball drawn is
red. The event B occurs if
Thus
P(B) depends on A1, A2, A3 which are mutually and exhaustive events. Therefore,
= P(A1) P(B Ι A1) + P(A2) P(B Ι A2) + P(A3) P(B Ι A3) = 3/10.2/9 + 2/10.3/9 +
5/10.3/9
Thus
Example 2.16
Suppose a factory has three machines M1, M2, M3 which produce 60%, 30% and 10% of
the total production respectively. Of their output, machine M1 produces 2% defective
items, machine M2 produce 3% defective items while machine M3 produces 4% defective
items. Find the probability that a part selected at random is defective.
51
Solution
Let B be the event “a part selected at random is defective”. A defective item could have
been produced by either machine M1, M2 or M3. Thus
Since (B ∩ M1), (B ∩ M2), (B ∩ M3) are mutually exclusive events. The following
information is contained din the equation.
Hence
=(0.6 × 0.02) + (0.3 × 0.03) + (0.1 × 0.04) = 0.002 + 0.009 + 0.004 = 0.025.
suppose you are now asked, what is the probability that a given defective part was
produced by machine M1. that is, you are to find P(M1|B) =P (a part was produced by
machine M1 given that the part was defective).
NOTE:
hence,
P( M i n B) P( M i ) P( B | M i )
P(M1|B) = 3
= 3
2.4
∑ P( M
i =1
i n B) ∑ P( M
i =1
i ) P( B | M i )
52
Thus,
Bayes Theorem
If A1, A2,…Ak are set of mutual exclusive and exhaustive events in a sample space Ω and
B is any other event in Ω such that P(B) > 0, then
P ( Ai ) P ( B | A i )
P(Ai|B) = k
; i = 1, 2,..., k 2.5
∑ P( A ) P ( B | A )
i =1
i i
Example 2.17
Suppose a college is composed of 70% male and 30% female students. It is known that
40% of the male students and 20% of the female students smoke cigarette. Find the
probability that a student observed smoking a cigarette is male?
Let M,F denote male and female respectively and S denotes smoker. The above problem
contains the following information.
P(S|M) = P(a student selected at random smokes given that the selected student is male) =
0.4
53
P(S|F) = P(a student selected at random smokes given that the selected student is female)
= 0.2
= P(A student selected at random is male given that the selected student is a
smoker)
P( M n S )
=
P( S )
Thus
Example 2.18
A table has drawers. Drawer 1 contains two red and five black biros, drawer II contains
four red and three black biros and drawer III contains one red and six black biros. A
drawer is chosen at random and a biro is chosen from the drawer. Find the probability
that
(ii) the biro chosen is from drawer I if the chosen biro is black.
Solution
54
Let P(i) denote the probability that drawer I is selected (I – 1, 2, 3) and R and B
representing red and black biros respectively.
Then
(ii) P(1|B) = P(the biro chosen came from drawer 1 given that the chosen biro is
black).
Definition 2.5
Let A, B and C be three events such that P(C) > 0. then the conditional probability of A
∪ B given C is defined by
Exercise 2.2
1. Suppose that a box contains 5 balls labeled 1 to 5. two balls are drawn at random
(one after the other without replacement).
(i) the sum of the numbers on the two balls selected is even?
55
(ii) The number on the first ball drawn is even if it is know that the sum of the
two numbers is even.
2. In a large population, it is observed that 30 percent of the people that are black
have caner and 25 percent of the people are not black have cancer. Assume that 10
percent of the population is black. What is the probability that a person selected at
random and found to have cancer is not black.
Suppose that, in a large population, 20 percent have been vaccinated. Find the
probability that a person who contracts smallpox has been vaccinated, assuming
that a vaccinated person without immunity has the same probability of contracting
smallpox as an unvaccinated person.
5. In a faculty of a certain college, 60% of the students are female: 20% of the
females and 50% of the male are studying mathematics. If a student data card is
selected at random and the student is found to be studying mathematics, what is
the probability that the selected student is a male?
6. In JAMB examination each question has 5 possible answers, exactly one of which
is correct. If a student knows the answer he selects the correct answer. Otherwise
he selects one answer at random from the 5 possible answers. Suppose that the
student knows the answer to 70% of the questions.
(i) What is the probability that on a given question the student gets the correct
answer?
(ii) If the student gets the correct answer to a question, what is the probability
that he knows the answer?
56
(iii) what is the expected score of the student in the examination?
7. A television set retailer finds out that 80% of this customers buy coloured T.V.,
and that 4 out of every 20 customers who buy coloured T.V set also buy antenna.
Calculate the probability that:
colored T.V
(iii) a randomly selected customer who has not bought an antenna has bought a
colored T.V set.
2.4 Independence
57
Therefore
Example 1.19
Toss a fair die twice and let A be the event “the first toss shows 3” and B be the
P(A) =
B ≡{(1,6), (6, 1), (2, 5), (3, 4), (4, 3), (5, 2)}
P(B) =
Therefore,
P(A ∩ B) =
Definition 2.6
58
(n-1) P(A1 ∩ A2 ∩ A3… An) = P(A1P(A2)P(A3)…P(An)
That is, the events A1, A2,…An are said to be mutually independent if
1 ≤ n1< n2…nk ≤ n.
Condition (i) is called pairwise independent. We might think that pairwise independence
always implies independence. But this is not necessarily so, as illustrated by the
following example.
Example 2.20
Let a pair of fair dice be rolled once. Consider the events. A1 number appearing on the
first die is even, A2 = the number appearing on the second die is odd = {1,3, 5} and A3 =
the difference of the two numbers is even
A1 ≡ {(2, 1), (2, 2), (2, 3), (2, 4), (2, 5), (2, 6), (4, 1), (4, 2),…, (6, 1), (6, 2),…}
A2 ≡ {(1, 1), (1, 3), (1, 5), (2, 1), (2, 3), (2, 5), (3, 1),…}
A3 ≡ {(1, 1), (1, 3), (1, 5), (2, 2), (2, 4), (2, 6), (3, 1), (3, 3), (3, 5), (4, 2), (4,4),
(4, 6), (5, 1), (5, 3), (5, 5), (6, 2), (6, 4), (5, 6)}
A1∩ A2≡ {(2, 1), (2, 3), (2, 5), (4, 1), (4, 3), (4, 5), (6, 1), (6, 3), (6, 5)}
A2∩ A3≡ {(1, 1), (1, 3), (1, 5), (3, 1), (3, 3), (3, 5), (5, 1), (5, 3), (5, 5)
A1∩ A2 ∩ A3 = Φ.
Thus
59
P(A1) = ½, P(A2) = ½, (P(A3) = ½
Since condition (ii) is not satisfied, we conclude that A1, A2 and A3 are not independent.
Definition 2.7
If the events A1, A2,…An are independent then
P(A1 ∩ A2… An) = P(A1)P(A2)…P(An) (2.8)
Example 2.21
A man fires 10 shots independently at a target. What is the probability that he hits the
target (i) 10times; (ii) at least once
If he has probability 1/3 of hitting the target on any given shot.
(i) Let Ai be the event “he hits the target at the ith shot” (I = 1, 2, 3,…10)
A1 ∩ A2 ∩ …∩ A10 is the event he hits the target 10 times A1, A2,… A10 are
independent events.
Therefore, the probability of hitting the target 10 times is
(ii) P(hitting the target at least once) = 1 – P(not hitting the target at all).
P(not hitting the target at all) = P(A1∩ A2 ∩ … ∩ A10)
Where, Ai ≡ not hitting the target at the ith shot. A1, A2, …, A10 are independent
events and P(Ai) = 1 - = .
Hence
Examples
60
2.22 Suppose a box contains 5 red and 3 black balls. A ball is chosen at random from
the box and then a second ball is drawn at random from the remaining balls in the
box. Find the probability that
(i) both balls are black
(ii) both balls are red
(iii) the first ball is black and the second is red
(iv) the second is red
(v) the second is black.
Solution
Let A1 be the event ”the first ball is red
Ā 1 be the event “the first ball is black
Ā 2 be the event ”the second ball is red
Ā 2 be the event “the second ball is black
(i) P(both balls are black) = P(Ā1 ∩ Ā 2)
P(Ā1 ∩ Ā 2) = P(Ā1)P(Ā 2)D Ā1)
P(Ā1) = P(Ā2DĀ1) =
Thus
P(Ā1 ∩ Ā 2) =
Thus
P(A1 ∩ A 2) =
(iii) P(A1 ∩ A 2) = P(first ball is black and the second ball is red) = P(Ā1)P(Ā 2) Ā1)
P(Ā1) = P(A2D Ā 1) =
Thus
P(Ā 1 ∩ A 2) =
61
(iv) P(A2) = P(A2 ∩ A1) + P(A2 ∩ Ā 1)
From (ii) and (iii) we have
2.23 Two fair dice are rolled. Given that the dice show different numbers, what is the
probability that at least one die shows a 6?
Solution
Let A be the event: the dice show different numbers.
A ≡ {(1,2), (2,1), (1, 3), (3, 1), (1, 4), (4, 1), (1,5), (5, 1), (1, 6), (6, 1),
(2, 3),…(5, 6), (6, 5)
B ≡ At least one die show a 6.
≡ {(1,6), (6,1), (2, 6), (6, 2), (3, 6), (6, 3), (4,6), (6, 4), (5, 6), (6, 5)}
A ∩ B = {(1,6), (6,1), (2, 6), (6, 2), (3, 6), (6, 3), (4,6), (6, 4), (5, 6), (6, 5)}
P(BDA) =
2.24 Let A and B be any two events defined on the same sample space. Suppose P(A) =
0.3 and P(A B) = 0.6. Find P(B) such that
(i) A and B are independent
(ii) A and B are mutually exclusive.
Solution
(i) If A and B are independent, then P(A ∩B) = P(A)P(B)
Thus
P(A B) = P(A) + P(B) – P(A)P(B)
We have
0.6 = 0.3 + P(B) – 0.3 P(B) = 0.3 + 0.7 P(B)
62
P(B) = = 0.43
hours half of the time and B one third. Find the probability that forany call during
the working hours
(i) no one is in to answer the call
(ii) A call can be answered by the person being called
(iii) Two successive calls are for the same woman
(iv) A caller who wants A has to try more than two times to get her.
Solution
(i) P(A and B are not in the office) =
P(X > 2) = 1 -
Exercise 2.3
63
1. (a) Show that if A and B are independent events, then
(i) A and Bc, (ii) Ac and Bc are also independent.
2. Let A and B denote two independent events such that A is a subset of B. prove that
either P(A) = 0 or P(B) = 1.
3. A man fires 10 shots independently at a target. The probability of hitting the target
at any shot is 1/3. Calculate the probability that
(i) noneof the shots hits the target (ii) At least one shot hits the target
(iii) The target is hit at least twice if it is know that it is hit at least once.
4. A box contains 6 red balls and 4 white balls. Three balls are drawn from the box
one after the other without replacement. Find the probability that
(i) the first two are white and the third red
(ii) the first two are white and the third white.
(iii) two are red and one is white
(iv) the second ball drawn is red
(v) the third ball drawn is white
5. A die is rolled 8 times. What is the probability that
(i) exactly 2 sixes appear.
(ii) at least 2 sixes appear.
(iii) at most 2 sixes appear.
6. Prove that if A1,…,An are independent events then
(i) P(A1 A2 … An) = 1 – [1 – P(A1)}{1 – P(A2)]…[1 – P(An)]
(ii) P(A1 A2 … An) ≤ 1- e –[P(A1) + P(A2) +…+ P(An)]
- (x + x +…+ x )
Hint: (1 – x1)(1 – x2)…(1 – xn) ≥ e 1 2 n xi ≤1
7. Show that P(A∩B∩C) = P(ADB∩C)P(BDC)P(C).
8. A die is tossed n times. What is the probability that a 6 appears at least two times
in the n tosses.
9. Suppose that A or B occurs is 0.7 while P(A) = 0.2, find P(B).
10. A boy decides to continue tossing a fair coin until he has thrown total of three
heads. Find Pn, the probability that exactly n tosses will needed.
64
11. Six blood samples are selected from 40 blood samples, of which four are
cancerous. What is the probability that exactly two of the blood samples selected
are cancerous?
12. Prove that if A1, A2,…An are any n events, then
P(A1 ∩ A2 … ∩ An) > 1 – {P(A2c) + P(A2c) + …+ P(Anc)}.
13. Prove that
(i) P(A ∩ Bc) = P(A) – P(A ∩ B)
(ii) P(A ∩ Bc) = 1 - P(A) – P(B) P(A ∩ B)
(iii) P(A) = P(A ∩ B) + P(A ∩ Bc)
(iv) P(A ∩ B) ≥ P(A) + P(B) – 1.
14. three coins have probability 0.5, 0.6 and 0.8 for heads respectively. One of then is
selected at random, that is, with equal chance for each, and tossed. If the outcome
is head, what is the probability that the coin with probability 0.8 for heads was
selected?
15. On a certain weekend there are 4 movies. Calculate the probability that at least one
of A and B will be selected by one or more of the 3 students.
16. A box contains n white balls numbered 1 to n, n black balls numbered 1 to n, and n
red balls numbered 1 to n. if two balls are drawn at random without replacement,
what is the probability that both balls will be of the same colour or bear the same
numbers.
17. On the first round, three fair coins are flipped at random. The coins resulting in
heads are flipped at random on the second round. If the second round results in
exactly one head, what is the conditional probability that the first round ended
inexactly two head?
65
UNIT THREE
DISCRETE RANDOM VARIABLES
3.1 Introduction
This chapter introduces idea of a random variable and its probability density function.
A random variable is a variable whose actual numerical value is determined by
chance. There are two easily identifiable types of random variables, discrete and
continuous. A discrete variable is one that takes only a limited number of possible
values, otherwise the variable is called continuous. This chapter is devoted to discrete
random variables.
Section 3.2-3.4 are devoted to some special discrete random variables. Bernoulli,
Binomial Poisson, uniform,Geometric and negative bionmial.
Example 3.1
Consider variable X, the number of heads in three tosses of a coin. There are four
possible values (0, 1,2, 3,) of X.
The actual value assumed is due to chance therefore X is a random variable. The
sample space for this experiment is
Ω= {HHH, HHT, HTH, THH, HTT, THT, TTH, TTT}
X = 0 if the outcome is TTT
X = 1 if the outcomes is HTT, or THT, or TTH
X = 2 if the outcome is HHT or HTH or THH
X= 3 if the outcome is HHH
Let p be the probability of the con landing tail. Since landing tail and landing head are
exhaustive events the probability of the coin landing head is 1-p.
P(T ∩ T ∩ T) = P(T)P(T)P(T)
Since the outcome at each trials are independent.
P(TTT) = p.p.p=P3
66
P(HTH) = P(T)P(H)P(H) = p (1-P) (1-P) = (1-P)2
P(HTH)= P(H)P(H)P(H)=( 1-P) (1-P) = P(1-P)2
P(HHT)= P(H)P(H)P(T) = (1-P)(1-P)P==(1-P)2
The probability of getting two heads = P (THH) + P(HTH) + P(HHT) = 3P (1-p)2
Similarly we have
P (o head) = P(TTT) = p3
P (1 head) = P(HTT) + P(TTH) + P(THT) = 3p2(1-p)
P (3 heads) = P(HHH) = (1-p)3
Thus we have the following table
Heads 0 1 2 3
Probability P3 3p2 (1- p) 3p (1 – p)2 (1 –p)3
Definition 3.1
A random variable X on a sample space Ω is a function assigns to each element Ω one
and only one real number X ( ) = x, the space ofX is the set of real number
Φ = {x:x= x ( ), we Ω).
Definition 3.2
A random variable x is discrete if t can assume at most a finite or a countable infinite
number of possible values.
In the above example,
Ω=
Where w1 = HHH, w2 = HHT…, w8 = TTT
67
= {0, 1, 2, 3} and X (w1) = 3, X (w2) = 2.X(w3) = 2, X (w4) = 2
X (w5 = 1, X (w6) = 1, X (w7 = 1, X (w8) = 0
That is {w:X(w) = x1} id an event
Definition 3.3
The real valued function f defined on R by f(x) = p(X = x) is called the discrete
probability density function of X.
Let X be a discrete random variable and suppose that the values it can assume are x1,
x2…,xn
The probability can be written as
P(X = x1) = f(x1), =P(X = x2) = f(x2)…, P(X = xn) = f(xn)
Such that
F(x) is calledprobability density function of X
For an illustration, let us consider the following examples.
Example 3.2
Suppose a pair of fair dice is tossed onc, Let X,Y,Z represent the sum, maximum and
minimum respectively of the two numbers appearing find the probability density function
of
(i) X, (ii) Y, (iii) Z.
Solution
The sample space Ω = {( 1, 1), (1, 2) ,…, (6, 6)} consist of 36 elements
(i) P (X =2) = P {( 1, 1)} = 1
36
P ( X = 3) = P{( 1, 2), (2, 1) = 2
36
P (X = 4) = P{(1,3), (3, 1) , (2, 2)} = 3
36
.
.
P (X = 12) = P {(6, 6)} = 1
36
Thus,
F(2)= 1, f(3) = 2, f(4) = 3 ---
68
36 36 36
In tabular form we have
X 2 3 4 5 6 7 8 9 10 11 12
F(x) 1 2 3 4 5 6 5 4 3 2 1
36 36 36 36 36 36 36 36 36 36 36
in function form
69
(iii) Z = Minimum of the two numbers
The possible values of Z are 1, 2, 3, 4,5,6
P(Z-1) = P{(1,6), (6,1)(1,5)(5,1)(1,4)(4,1)
(1,3)(3,1)(1,2)(2,1)(1,1)! = 11
36
P(Z = 2) = P (2,6)(6,2)(2,5)(5,2)(2,4(4,2)
(2,3)(3,2) (2,2)= 9/36
P(Z= 4)P{(4,6) (6,4) (4, 5) (5,4) (4,4) 5.36
P(Z =5) = P(5,6) (6,4) (5,5) 3/36
P(Z= 6) = P(6,6) 1/36
Putting in form of a table we have
Z 1 2 3 4 5 6
h(z) 11 9 7 5 3 1
36 36 36 36 36 36
H(Z) = 13 – 2z z = 1, 2, 3, 4,5, 6
36
= 0 for other values of x.
The probability density function of a discrete random variable X has the following
properties
(i) O < f(x) < 1, x E R
(ii) {X:f (x) = 0} is a finite or countable infinite subset of R
(iii) ∑f(xi) = 1
The above densities can be represented in terms of a diagram as illustrated in figure (a),
(b) (c)
70
Example 3.3
Let X,Y,Z be the random viable introduced in example 3.1 above
The above three properties are satisfied, properties (i) & (ii) are immediate from
definition of probabilities. To check (iii) we have
∑f(x1) = 1/36+ 2/36+ 3/36 + 4/36 + 5/36+ 6/36 +5/36 + 4/36 + 3/36 + 2/36+ 1/36 =1
Example 3.4
Suppose a box contains 1 balls of which 4 are red and 6 are black. A random sample of
size 3 is selected. Let X denote the number of red balls selected. Find the probability
density function of x if
10c5 9 8 6
P(X = 3) = 4C3/10C3 = 1
71
30
P(X = 0) =P (first ball is black, second black and the third black) = P(bbb).
6/10= 3/5
P(bbb) =
Similarly ,
P(X = 2) =
P(X = 3) = P(RRR) =
X 0 1 2 3
f(x)
72
This can be written as
f(x) =
0 elsewhere
Example3.5
A box contains 6 balls labeled 1, 2, 3, 4, 5, 6 two balls are drawn at random one after
the other. Let X denote the larger of the two numbers on the balls selected, obtain the
probability density function of X if
(i) Sampling is without replacement (ii) Sampling is with replacement
Solution
(i) The possible values of X are 2, 3, 4, 5, 6. The larger of the two numbers can not be
1 since sampling is without replacement and we can not get same number twice.
P(X =5) = P{(1, 5), (5, 1), (2, 5), (5, 2), (3, 5), (5, 3), (5, 4), (4, 5)} =
Similarly,
P(X = 6) =
X 2 3 4 5 6
73
f(x) 1/15 2/15 1/5 4/15 1/3
f(x) = x = 2, 3, 4, 5
0 elsewhere
(ii) The possible values of X are 1, 2, 3, 4, 5, 6. The larger can be 1 in this case
since (1, 1) is a possible outcome. The sample space consists 36 possible
outcomes
P(X = 1) = P(1, 1) =
P(X = 6) = P{(1, 6), (6, 1), (2, 6), (6, 2).,,, (6, 6)} =
X 1 2 3 4 5 6
f(x) = x = 1, 2, 3, 4, 5, 6
3.6 A couple decides that they will continue to have children until either they
have a boy and a girl in the family or they have four children. Assuming that boys
and girls are equally likely to be born. Let X denote the number of children in the
family. Find the probability density function of X.
74
That is, the couple will stop having more children if the first child is a boy and the
second of girl or the first is a girl and the second a boy.
Thus,
Similarly,
X 2 3 4
f(x) ½ ¼ ¼
Exercise 3.1
1. A fair coin is tossed until a head or five tails occur. Let X denote the number of
tosses of the coin. Compute the probability density function of X.
2. A box contains 2 red balls and 3 blue balls. Balls are successively drawn without
replacement until a blue ball is drawn. Let X denote the number of draws required.
Compute the p.d.f of X
3. The pdf of a random variable X is given by
X 1 3 4 6 8
f(x) K
75
4. Toss a fair coin two times. Let X be the number of heads obtained. Find the pdf of
X.
5. A coin with probability p of a head is tossed until a head appears. Let X denote the
number of times the coin is tossed. Find the pdf of X.
6. A fair die is tossed twice. Let X denote the product of the two numbers appearing.
Find the pdf of X
7. A fair coin is tossed3 times. Let X represent the difference between the number of
heads and the number of tails obtained. Find the pdf of X.
8. The pdf of a random variable X is given by
f(x) = {k 2x, x = 1, 2, 3, …N, zero elsewhere}. Find the value ofK.
Definition 3.3
Probability Distribution function. Let X be a random variable with probability density
function f(x). The probability distribution function of X denoted f(x) is defined by
F(x) = p(X ≤ x) for x real.
=
Properties of the Probability distribution function
1. F is a non decreasing function, that is if a < b, then
F(a) < F(b).
2. lim F(b) = 1.
b-∞
lim F(b) = 0
b-∞
3. F is right continuous. That is F(b + ) = F(b).
Let A be any subset of R and let f(x) be the probability density function of X. ew can
compute P(X ε A) by noting that {w : X (w) ε A} is an event and that
{w : X(w) ε A} = U { w : X(w) = xi}.
XiεA
76
Thus
P(X ε A) =
If A is an interval with end points a and b, say A = [a,b].
Then
{(X ε A) = P(a ≤ X ≤b) =
Example 3.7
Consider the random variable X the sum of the two numbers appearing when a fair die is
tossed twice of Example 3.2. The probability distribution function for X is given by
X 2 3 4 5 6 7 8 9 10 11 12
F(x)
Suppose we wish to find the probability that X is between 4 and 9 (4 and 9 inclusive) we
write it as
P(4 ≤ X ≤ 9) = P(X = 4) + P(X = 5) + P(X = 6) + P(X = 7) + P(X = 8) + P(X = 9)
=
coin is fair and let denote the outcome of the toss. Then there are two possible values for
X, Heads or tails. These two values are mutually, exclusive and exhaustive and we may
associate the two possible outcomes of the toss with values 1, 0 of the random variables
X. That X = 1 when a head appears and X = 0 when a tail appears.
P(X = 1) = p, P(X = 0) = 1 –p.
The p.d.f. of X is
X 0 1
f(x) 1–p P
Or in functional form
F(x) = pX(1 –p)1-x, x = 0, 1
0 elsewhere 3.1
f(x) as defined above is called the Bernoulli probability density function and any variable
X having (3.1) has its probability density function is called a Bernoulli random variable
and is said to have the Bernoulli distribution.
78
variables X1, X2,…Xn are independent Bernoulli random variables. Let us assume the
probability of success is p and failure 1-p and
P(x1 = 1) = p
Then, sum Sn = + x2 +---+ Xn is the number of successes in n Bemoulli trials. That is, Sn
is a counting- variable counting the number of successes in n repeated trials. This random
variable Sn is called the Binomial random variable. The possible values of Sn are 0, 1, 2,
3, 4,----, n.
P(Sn = 0) = p(no success)
= p(1sttrail is a failure) P(2nd trail is a failure)…P(nth trail is a failure).
= P(X1 = 0) P(X2 = 0) P(X3 = 0)…P(Xn = 0)
= (! – p) (1 – p)(1 – p)…(1 –p) = (1-p)n.
Sn is 1 if the sequence of outcome is
1 0 0 0 0 0 0 …0 or 0 1 0 …0 or 0.0 1 0 …0,…0 0 0 0 0 0 1
P(Sn = 1) = P(1, 0, 0, 0,…0) + P(), 1, 0, 0,…, 0) + …+ P(0, 0,…1)
= P(X1 = 1)P(X2 = 0)…P(Xn = 0) + P(X1 = 0)P(X2 = 1)…P(Xn = 0) +…+
+ P(X1 = 0)…P(Xn = 1)
= P(1 – p)…(1 – p) + (1- 1)P(1 –p)…(1 – p)…(1 – p)+ (1 – p)(1 – p)
= P(1-p)n-1 + P(1-p)n-1 +…+ P(1-p)n-1 = np(1-p)n-1
Similarly,
P(Sn = 2) = nC2 P2 (1-p)n-2
Where nC2 is the number of sequences in which exactly 2 have value 1 and the others 0
e.g. (1, 1, 0, 0, 0, 0,0,…, 0), (1, 0, 1, 0,…, 0)…
In general, it can easily be seen that
P(Sn = k) = nC2 pk (1-p)n-k, k = 0, 1,…, n
Where nC2 is the number of sequences in which exactly k have value 1 and others 0.
For example when n = 4, possible sequences of outcomes are given below.
Sequence Sn P(Sn)
79
(0,0,0,0) 0 (1 – p)4
(1,1,1,0) 3 P3(1 – p)
(1,1,0,1) 3 P3(1 – p)
(1,0,1,1) 3 P3(1 – p)
(0,1,1,1) 3 P3(1 – p)
(1,1,1,1) 4 P4
Thus,
P(Sn = 0) = (1 – p)4
P(Sn = 1) = 4p(1 – p)3
P(Sn = 2) = 6p2(1 – p)2
P(Sn = 3) = 4p3(1- p)
P(Sn = 4) = p4
Theorem 3.1
Let Sn denote number of successes in n repeated Bernoulli trails, with probability of
success p. the probability density function of Sn is given by
80
f(x) = P(Sn = x) = nCx px(1 – p)n-x x = 0, 1, …,n (3.2)
0
Definition 3.5
A discrete random variable X denoting total number of successes in n trails is said to
have the binomial distribution if
Example 3.8
A soldier fires 10 independently at a target. Find the probability that he hits the target.
(i) once (ii) at least 9 times (iii) at most two times.
If he has probability 0.8 of hitting the target at any given time? Let X denote the number
of times he hits the target. Then X is a binomial variable with n =10 and p = 0.8
From equation (3.2), we have
P(X = x) = 10Cx (0.8)x (0.2)10-x
(i) P(X = 1) = 10Cx 0.8(0.2)9 = 8(0.2)9
(ii) P(He hits the target at least 9 times) = P(X ≥ 9)
= P(X = 9) + P(X = 10)
P(X = 9) = C9 (0.8)9 (0.2) = 10 × (0.8)9 × 0.2 = 2(0.8)9
10
Hence,
P(X ≥ 9) = 2(0.8)9 + (0.8)10 = (0.8)9(2 +0.8) = (0.8)9(2.8) = 0.3758
(iii) P(at most twice) = P(X ≤ 2) = P(X = 1) + P(X = 2)
P(X = 0)= (0.2)10;
P(X = 1) = 8(0.2)9;
P(X = 2) = 45 (0.8)2 (0.2)8
Thus,
P(X ≤ 2) = (0.2)10 + 8(0.2)9 + 45(0.8)2 (0.2)8 = 0.00008.
Example 3.9
A fair die is rolled four times. Find the probability of getting 2 sixes. Let us call a six a
success on a toss of a die and let X be number of sixes (successes) in 4 trails. X is
binomial random variable with n = 4 and p = 1/6. Thus,
81
P(X = 2) = 4C2 ..
Example 3.10
Suppose that a certain type of electric bulb has a probability of 0.3 of functioning more
than 800hours. Out of 50 bulbs, what is probability that less than 3 will function more
than 800 hours. Let X be the number of bulbs functioning more than 800 hours.
Assuming that X has a Binomial distribution,
P(X = x) = 50 Cx(0.3)x(0.7)50-x
P(X < 3) = P(X = 0) + P(X = 1) + P(X = 2)
P(X = 0) = (0.7)50; P(X = 1) = 50.(0.3) (0.7)49
P(X = 2) = 50C2 (0.3)2(0.7)48.
Thus,
P(X < 3) = (0.7)50 + 15(0.7)49 + 110.25(0.7)48 = 0.0000046
Exercise 3.2
1. A fair coin is tossed 4 times. Compute the probability that (i) exactly two heads
occur (ii) at least 3 heads occur.
2. An investigation reveals that four out of every five patients are cured of malaria
when treated with a new drug. If a sample of ten patients is treated by the new
drug, compute the probability that (i) exactly six patients are cured, (ii) at most
four patients are cured
3. A man fires 12 shots independently at a target. The probability of hitting his target
is ¼.
(i) What is the probability of his hitting the target at least two times?
(ii) How many times must he fire so that the probability of his hitting the target
at least once is greater than 7/9?
4. Six children are born in a hospital in a given day. Calculate
(i) the probability that the number of boys is the same as the number of girls.
(ii) the probability that there are more girls tha boys
(iii) the most likely number of boys.
82
5. An experiment has 90 percent probability of success and 10 percent probability of
failure. The experiment is repeated four times. Find the probability of obtaining
(i) No success (ii) No failure (iii) two successes and two failures.
6. Let X be a Binomial random variable with paraments (n.p). show that
(i) P(X = k) first increases monotonically and then decreases monotonically.
(ii) The value of k that maximizes P(X = k) is the largest integer less than or
equal to (n + 1)P
(iii) P(X = k) =
7. A fair coin is tossed repeatedly until three are obtained. Find Pn, the probability
that exactly n tosses are needed.
n
8. Show that (i) Cxpx (1 – p)n-x = 1 (ii) n
Cxpx (1 – p)n-x = np
0<p<1.
P(X = x) = 0 otherwise
Where X is the number of trials before the first success and if we define Y as the number
of failures preceding the first success, we have
83
P(1 – p)y y = 0,1,2…
P(Y= y) = 0 otherwise
Definition 3.6
A discrete random variable X is called a geometric random variable if its probability
density function is given by
Where X is the number of independent Bernouilli trials taken for the first success to
occur (the successful trial is included in the count). To see that (f(x) is a probability
density function, all that needed to be checked is that
(1-p)x=1= 1 = P[1 + (1-p) + (1-p)2 +…]
From geometric series,
(1-p)x=1= .
Example 3.11
A fair is tossed until a head appears (a) What is the probability that three tosses are
needed? (b) What is the probability that at most three tosses are needed?
Solution
Let X denote the number of tosses until a success (a head) occurs. Since the coin is fair,
.
P(X = 3) = (1 – p)3-1
P(X = 4) = . .
5 1
6 6
Thus
2 3
P(X ≤ 4) = + . + . + .
1 5 1 5 1 5 1
6 6 6 6 6 6 6
x =1
1 4 5 1 1 − (5 / 6) 4
= ∑ = . = 1 − (5 / 6) 4
6 x =1 6 6 1− 5 / 6
or
11 25
P(X ≥ 3) = 1 – {P(X = 1) + P(X = 2)} = 1 - =
36 36
Example 3.13
The probability that a certain test yields a “positive” reaction is 0.6. what is the
probability that not more than 4 negative reactions occur before the first positive one?
Solution
85
Let X denote the number of negative reactions before the first positive one, then
P(X = x) = (0.4)x(0.6); x = 0,1, 2,…
Thus,
4 4 4
P(X ≤ 4) = ∑ P( X = x) = ∑ (0.4) X (0.6) = (0.6) ∑ (0.4) x = (1 − (0.4) 5 ) = 0.9898
0 x =0 x=0
The geometric random variable has an interesting property which is summarized in the
following theorem.
Theorem: 3.2
Suppose that X is a geometric random variable. Then for any given two positive integers
s and t,
P(X > s + t|X > s) = P(X > t).
Proof:
P ( X > s + t and X > s ) P ( X > s + t
P(X > s + t | X > s) = =
P( X > s) P( X > s)
∞
P(1 − p ) s
P(X > s + t) = P ∑ (1 − p)
S +1
X −1
=
1 − (1 − p )
= (1 − p ) s
Thus,
(1 − p ) s +t
P(X > s + t| X > s) = = (1 − p) t
(1 − p ) s
∞
p (1 − p) t
P(X > t) = ∑ p(1 − p) X −1=
t +1 p
= (1 − p) t .
Hence,
P(X > s + t |X > s) = P(X > t).
The above theorem states that if a success has not occurred during the first s repetitions of
Bernouli trails, then the probability that it will not occur during the next t repetitions is
the same as the probability that it will not occur during the first t repetitions of Bernoulli
trails. Therefore the distribution is said to have “no memory”.
86
Exercise 3.3
1. On a certain road the probability of an accident on any day is 0.05 (assuming not
more than one accident can occur on any day). Assuming independence of
accidents from day to day on this road, what is the probability that the first
accident of the year occurs in the month of March.
2. A fair die is rolled until a six appears. Calculate the probability that he has to
throw the die more than three times before he gets a six.
3. Let X be a geometric random variable with p = 0.2. Calculate the following
probabilities. (i) P(3 < x ≤ 6) (ii) P(2≤ X≤ 4) (iii) P(X ≤ 2).
∞
1
4. Prove that ∑ xp(1 − p)
x =1
x =1
=
p
.
Example 3.15
A fair die is rolled until two sixes occurs, find the probability that
(i) exactly 5 tosses are needed
(ii) at most 5 tosses are needed.
Solution
Let X be the number of tosses needed to get two sixes. X is a Pascal random variable
with pdf: f(x) = x-1C1p2(1-p)x-2, x = 2,3,..
Where p = ½ , r = 1.
(i)
(ii)
Thus,
88
3.3.2 The Hyper geometric Random Variable
Suppose that we have a box containing n items of which n1 are defective and n – n1 are
non-defective. Suppose that we choose at random k items from the box without
replacement. Let X denote the number of defective items in the k items selected. Then
(3.4)
The reader will notice that X = x if and only if we obtain x defective items from n1
defective items in the box (n1Cx ways of doing this).
Any random variable having its probability density function as given by (3.5) is called
hyper geometric random variable and is said to have hyper geometric distribution.
Now as n ∞.
1.
2.
3.
Consequently,
n
Cxpx(1 – p)n-x =
(3.5) gives an approximation to the binomial distribution with λ = np, when n is large and
p is small, where e = 2.71828 is the base of natural logarithms.
Definition:
A random variable X is called a Poisson random variable if its probability density
function is given by
89
(λ > 0) (3.8)
0 otherwise
X represents the total number of events which have occurred up to time t. When t= 1 then
f(x) corresponds to the probability density function of number of events in a unit interval.
For examples, the number of calls that come into a telephone exchange in a unit time
interval, the number of vehicles passing through a designated point in a unit time interval.
In order to motivate our discussion, let us consider the following examples.
Example 3.16
Suppose a rare disease occurs in 2 percent of a large population. A random sample of
10,000 people are chosen at random from this population and tested for the disease.
Calculate the probability that at least two people have the rate disease.
Solution
The probability p of having the disease is 0.02 and n = 10,000. Let X be the number of
people having the disease. X is a binomial random variable with parameters n= 10,000
and p = 0.02. we shall apply the result (3.5) since n is large and p very small.
Hence,
Np = 0.02 × 10,000 = 20,,
Thus,
P(X = x) =
P(X = ) = e-200
P(X = 1) = 200 e-200
Hence
P(X ≥ 2) = 1 – P(X < 2) = 1 – e-200 (1 + 200) = 1 – 201 e-200.
Example 3.17
On a given road, an average of five accidents occur every month. Calculate the
probability that over a year period there will be (i) no accident (ii) at most 2 accidents.
Solution
In this case, λ = 5, t = 12. From (3.8) we have
90
= e-60 (1 + 60 + 1800) = 1861 e-60.
Tables for the Poisson distribution are available and brief abulation is given in the
Appendix.
Exercises 3.4
1. If 3% of the items manufactured in a factory are defective. Compute the
probability that in a simple of 100 items. (i) 2 items will be defective (ii) at least
2 items will be defective. Use Poisson approximation to the Binomial.
2. Use the Poisson approximation to calculate the probability that at least two sixes
are obtained when six dice are rolled once.
3. The telephone switchboard of a University has an average of two incoming calls
per minutes. Calculate the probability that, over a three-minute interval, there will
be (i) no incoming calls (ii) exactly one incoming call. (iii) at most two incoming
calls.
4. Let Pr be the negative binomial pdf with parameters r and P. prove that
where Pr(k) = P(X= k).
91
13. If n ≥ 8 and in s binomial distribution the probability of 7 successes in n trails is
equal to the probability of 8 successes in n trail, find the probability of success on
any given trail.
14. Let X be a random variable with moment generating function MX(t). If R(t) = In’’
MX(t) for all t. Find R”(0).
15. If a fair die is rolled repeatedly, find the probability that the first “six” will appear
on an odd-numbered roll.
p, x = -1
16. f(x) =
(1-p)2px, X = 0,1,2
(i) Show that f(x) is a pdf
UNIT FOUR
EXPECTATION OF DISCRETE RANDOM VARIABLE
In this chapter we introduce the concept of the mean value of a random variable. It is
closely related to the notion of weighted average. The moment and probability generating
functions of a random variable are also introduced.
4.1
Let X be a random variable having possible values x1, x2, …, xn, and let the experiment
on X be performed n times. For example let X be the outcome of rolling a die. There are
six possible values x1 = 1, x2 = 2, x3 = 3, x4 = 4. x5 = 5 and x6 = 6. Suppose the die is
rolled n times. The successive rolls constitute independent repetitions of the same
experiment. Let X1, X2, . …, Xn denote the outcomes of the experiment of rolling a die n
times (that is Xi denote the outcome of the ith toss). Then
Is the average of the numbers that appeared. Let fi denote the number of times xi occur.
Then we have
We know that
is
called the expectation of the random variable X. thus the expected value of a random
variable X is the long-run theoretical average value of X.
92
Definition 1
Mathematical Expectation. Let X be a random variable with probability density function
as follows:
Value of X, x x1 x2 x3 …. xk
Examples
4.1 A fair coin is tossed three times. Let X denote the number of heads obtained. Find
the mathematical expectation of X.
Solution
The probability density function of X is as follows.
X 0 1 2 3
f(x)
Hence,
E(X) = (1 × 1/8) + (1 × 3/8) + (2 × 3/8) + (3 × 1/8)
= 0 + 3/8 + 6/8 +3/8 = 12/8 = 1.5
4.2 What ate the mathematical expectations of the variables X, Y and Z as defined in
example 3.1.1 of chapter 3.
Solution
The probability density function of X is
X 2 3 4 5 6 7 8 9 10 11 12
f(x) 1/36 2/36 3/36 4/36 5/36 6/36 5/36 4/36 3/36 2/36 1/36
The p.d.f of Y is
Y 1 2 3 4 5 6
The p.d.f of Z is
Z 1 2 3 4 5 6
Exercises 4.1
1. Suppose a box contains 10 balls of which 4 are red and 6 are black. A random
sample of size 3 is selected. Let X denote the number of red balls selected. Find
(E(X) if
(i) Sampling is with replacement.
(ii) Sampling is without replacement.
2. A box contains 6 balls labeled 1, 2, 3, 4, 5, 6. Two balls are drawn at random one
after the other. Let X denote the larger of the two numbers on the balls selected.
Compute E(X) if (i) Sampling is with replacement, (ii) Sampling is without
replacement.
3. A box contains 3 balls and 2 white balls. A ball is without replacement one after
the other until a white ball is drawn. Find the Expected number of draws required.
94
since p + 1 – p = 1. Hence,
E(X) = np.
Example 4.3
A fair die is rolled 12 times, what is the expected number of six appear?
Solution
Let X be the number of sixes that appear. X is a binomial random variable with n = 12
and P = 1/6 (probability of a six). Hence
E(X) = np = 12x 1/6 =2.
That is, we expected to get 2 sixes when a die is rolled 12 times.
Definition 2
Let X be a discrete random variable having probability density function f(x).
Examples
4.4 The Random Variables X has Probability
X -2 -1 0 3 5
Solution
(ii)
= -2/4 – 1/8 +3/4 +5/4 = 11/8 = 1
(iii) The possible values of 3x are -6, -3,0, 9, 15.
The probability density function of 3x is
3x -6 -3 0 9 15
95
E(3X) = (-6 × ¼) + (-3 × 1/8) + (0 × 1/8) + (9 ×1/4) + (5 × ¼)
= -6/4 – 3/8 + 9/4 + 15/4 = 33/8
Note that
P(3X = -6) = P(X = -2_,
And so on and
E(3X) = 33/8 = 3 × 11/8 = 3E(X)
(iii)
x+5 3 4 5 8 10
(iv)
X2 4 1 0 9 25
Definition 3
Let X be a random variable whose probability density function is given by
x x1 x2 …………. xk
Let ψ(X) be a function of X. then the mean or expected value of the new random ψ(X) is
given by
E[ψ(X)] = ψ(x1) f(x1) + ψ(x2) f(x2) +…+ψ(xk) f(xk)
That is
E[ψ(X)] = ψ(xi) f(xi)
Theorem 1
Let X be a random variable and let ψ(X) = a X + b where a, b are constants, then
96
E(ψ(X)) = aE(X)+ b
Proof:
Suppose the probability density function of X is
{xi, f(Xi), I = 1, 2, …,k)
Properties of Excitations
1. If c is a constant and P(X = c) = 1, then E(X) = c.
2. If a and b are constants and X has a finite expectation, then aX has finite
expectation and
(i) E(aX) = aE(X) + b.
(ii) E(aX + b) = aE(X) + b.
3. If X and Y are two random variables having finite expectations, then
(i) X + Y has random finite expectation and E(X + Y) = E(X) + E(Y)
(ii) E(X) ≥ E(Y) if P(X ≥ Y) = 1.
(iii) DEXD ≤ EDXD.
4. A bounded random variable has a finite expectation. That is if P(X ≤M) = 1, then
X has a finite expectation and DEXD ≤ M.
Definition 4
Variable. Let X be a random variable with mean E(X) = µ.
The variance of X, denoted by Var(X) is defined by
Ver(X) = E{X - µ)2} =
A general formula that is usually simple for computing the variable is given below
Var(X) = E{(X - µ)2} = E{X2 – 2 Xµ + µ 2}
= E(X2) – E(2µX) + E(µ 2) = E(X2) - 2µE(X) + µ 2 (property 2)
= E(X2) - 2µ 2 + µ 2 = E(X2) - µ 2.
Thus, the computing definition of variance is
Var(X) = E(X2) – [E(X)]2 (7)
The variance of X is interpreted as a numerical measure of spread or dispersion about its
mean.
Example 4.5
1. Find the variance of a random variable having the following probability density
function.
x 1 2 3 4 5 6
97
f(x) 1/6 1/6 1/6 1/6 1/6 1/6
E(X2) = (12 × 1/6) + (22 × 1/6) + (32 × 1/6) + (42 × 1/6) + (52 × 1/6) + (62 × 1/6)
= 1/6 + 4/6 + 9/16 + 16/6 + 25/6 + 36/6 = 91/6.
Hence from (7) we have
Var(X) = 91/6 – (21/6)2 = 105/36 = 2 (11/12).
Example 4.6
Find the variance of a Bernoulli random variable with parameter p. that is X has
the following p.d.f.
X 0 1
f(x) 1–p P
E(X) – 0 × (1-p) + (1 × P) = p
E(X2) = 02 × (1-p) + (12 × p) = p.
Hence,
Var(X) = p - p2 = p(1 – p)
Properties of Variance
1. If c is a constant and P(X) = c) = 1, then Var (X) = 0.
2. If a and b are contents, then
(i) Var(aX) = a2 Var(X)
(ii) Var(aX + b) = a2 Var(X)
(iii) Var(X) ≥ 0
The proof of property (1) is very trivial and this is left as an exercise to the reader,
we shall now give a proof of (ii). From (7),
Var(aX) = E[aX2] – [E(aX)]2 = E a2 X2 – [aE(X)]2
= a2 E(X2) – a2[E(X)]2 = a2{(E(X2)]2} = a2 Var(X).
Definition 5
Standard deviation. Let X be a random variable with mean µ.
The standard deviation is the positive square root of the variance, and is given by
98
To calculate the variable of X, we need E(X) and E(X2).
Var(X) = (E(X))2.
E(X) =np.
n
E[X(X – 1)] = (xpx(1 – p)n-x = px(1 – p)n –x
(w2hen x = 1 or 1 the expression is zero)
= px(1 –p)n-x
E[X(X – 1)] =
= p[2.1(1-p) + 3.2 (1 – p)2 +…+ r(r -1)(1 – p)r-1+…].
Multiply both sides by 1-p,
(1 – p) E[X(X- 1)] = p(2.1(1 –p)2 + (3.2(1 – p)3 +…+ r(r-1)(1-p)r +…+]
E(X(X-1)] – (1-p)E[X(X-1)] = p[2(1-p) + 4(1-p)2 + 6(1 – p)3 +…+ 2r(1 – p)r +…]
Hence
99
E[X(X)] – (1-p) E[X(X – 1)] =
E[X(X – 1)] =
E(X2) = + E(X) =
Thus
Var(X) =
E(X) =
= λe-λ e-λ eλ =λ
E[X(X – 1)] =
= e- λ λ2
E(X2) – E(X) = λ2
E(X2) = λ2 + E(X) =λ2 + λ
Hence,
Var(X) = λ2 + λ – λ2 = λ.
This shows that the mean and variance of a passion random variable are both equal.
Exercise 4.2
1. Compute the variance and standard deviation of the random variable X defined in
exercises 4.1.1
2. A fair coin is tossed twice. Let X denote the number of heads that appear.
Compute
(i) the expected value of X
(ii) the variance of X
100
(iii) the expected value of
3. Calculate the mean, variance and standard deviation of the random variable having
the following probability density function.
X -2 1 0 1 2
101
Definition 6
The probability generating function of a non-negative, integer-valued random variable X
is defined by
E(Sx) = P(S) = (8)
From (6), we see that
P(S) = E(Sx)
By differentiating with respect to s we have
P’ (s) = E(XSX-1)
P’’(s) = E[X(X – 1) SX-2]
P(r) (s) = E(X(X – 1)(X – 2)…(X – r + 1)sX-1)
Putting s =1 in the above derivatives, we have
P’(1) = E(X)
P’’(1) = E[X(X – 1)] = E(X2) – E(X)
P(r) (1) = E[X(X – 1)(X – 2)…(X – + 1r)].
Thus the mean and variance of X can be obtained from P(s) by the following formulas.
E(X) = P’(1)
E(X2) – E(X) = P’’(!)
E(X2) = P’’(1) + E(X) = P’’(1) + P’(1)
And
Var(X) = E(X2) – [E(X)]2 = P’’(1) + P’(1) – [P’(1)]2
We now illustrate the use of these formulas with the following examples.
Examples 4.7
Let X be a poisson random variable wit parameter λ. Find the mean variance of X
f(x) =
102
Thus,
P(S) = [sp + (1 – p)]n (12)
Differentiating, we have
P’(s) =np [sp + (1 – p)]n-1
P’’(s) = n(n – 1)p2 [sp + 1 – p]n-2
Var(X) = p’’ (1) + - [p’(1) – [p’(1)]2 = n(n – 1) p2 + np – (np)2 = np(1 – p).
The table below summarizes some of the results of this chapter
x = 0, 1, 2,
….n
x = 1,2,…
Uniform
X= a, a +
1,…b
RtX = 1 + tX + +…
E(etX) = E(1 + tX + + …+
Under fairly general conditions we shall assume that expectation of infinite sum equals
the sum of the expected value but this is true in general for finite sum
1 + tE(X) + + +…
That is
MX(t) = 1 + tE(X) +
M’’x(0) = E(X2+)
hence,
Var(X) = Mx’’(0) – [MX’(0)2]
And
Mx(t) = E(etX), Mx(r)(t) = E[Xr e tX] (15)
Putting t = 0, we have
M(r)(0) =E(Xr)
that is, E(Xr) is the rth derivative of Mx(t) evaluated at t=0. (16)
Examples 4.9
Let X be a binomial random variable with parameters n and p. then
MX(t) = E(etX) = n
CxPx(1 – p)n-x = n
Cx(Pet)x(1 – p)n-x
from the binomial theorem
(a + b)n = n
Cxaxbn-x = [pet + (1 – p)]n,
104
on putting a = pet, b = 1- p.
Example 4.10
Let X be a Poisson random variable with parameter λ. Then
Differentiating, we have
M’X(0) = λete-λ(1-et)
M’x(0) = λete-λ(1-et)Dλ=0 = λ E(X)
Mx’’(t) = λet e-λ(1-et) + λ2 e2te-λ(1-et)
Mx’’(0) = λ + λ2 = E(X2)
Hence,
Var(x) = λ + λ2 – λ2 = λ.
Example 4.11
Let X be a geometric random variable with parameter p. them
Mx(t) =
The fourth equality follows from sun to infinity of geometric series Further
applications of Mx(t) will be considered later in Chapter 6.
Exercise 4.3
1. Find the probability generating function of the geometric distribution with
parameter P. hence determine the mean and variance of the distribution.
2. Let X be uniformly distributed on (0, 1, 2,…n). Find the mean and variance of X.
Hint:
3. Let X be a random variable with finite variance. Prove that for any real number a,
Var(X) = E[X – a)2] – [E(X) + a]2.
4. A fair die is tossed 50 times. Let X be the number oftimes six appears. Evaluate
105
(i) E(X) (ii) E(X2) (iii) VAR(X)
5. Find the expected value and variance of the random variable X os Example 3.1.1.
6. Show that if c is a constant, Var(x + c) = Var(X).
7. Let X be uniformly distributed on (0, 1, 2,..N). Find P(S), mean and variance of X.
8. Let X be defined by
x 1 2 3
1 1 1
2 4 4
f(x)
10. Derive the MGF for the following: (a) Bernoulli, (b) Negative Binomial, (c)
Binomial and use it to find the mean and variance.
c(x + 2), x = 1, 2, 3, 4
f(x) =
0 elsewhere
Find (i) c, (ii) moment generating function of X, (iii) use result of (ii) tofind the
mean and variance of X.
12. Prove that for any random variable, X, E(X2) > [E(X)]2.
14. A fair die is rolled until all the 6 sides appeared at least once. Find the expected
number of rolls needed.
p, x = -1
15. Let f(x) =
Px(1-p)2, x = 0,1, 2,…
(a) Show that f(x) is a pdf. (b) Find Mx(t), the moment generating function of
(b) Hence, the mean and variance of X.
106
MODULE THREE (UNIT 1)
CONTINUES RANDOM
5.1 INTRODUCTION
In the chapter so far, we consider discrete random variables and their probability
density functions e.g. Binomial, Poisson, Geometric. These are essentially counting
variables. However, there exist random variables whose set of possible uncountable,
such random variables are called continuous random variables.
Example 5.1
Let a point be chosen at random from the interval (0, 2) and let X represent the point
chosen. Then 0 ≤ X≤ 2. the range space of X consists of unaccountable real numbers. Any
real number in the interval (0, 2) is a sample point. Let A be the event that the point
chosen is 0.3. From the basic definition of probability we have
107
number of po int A 1
P ( A) = = = 0.
number of po int s in Ω ∞
Hence,
P(a ≤ X ≤ b).
P(X ε A) = ∫A
f ( x)dx
Definition
108
A probability distribution function is any function F satisfying the following conditions:
(iv) F(x +) = F(x) for all values of x where F(x+) = Lim F(x + δ).
δ−0
Example 2
Show that
e x x<0
F ( x) =
1 x >0
Solution
109
0 ≤ ex ≤ 1 for x ≤ 0 and ex is a non decreasing continuous function, hence F(x) is a
probability distribution function.
x
F(x) = P(X ≤ x) = ∑ f ( y)
−∞
Similarly, for continuous random variables, there exists a function f(x) such that
x
F(x) = Px(X ≤ x) = ∫−∞
f ( y )dy
Such a function f is called probability density function of X. to see this consider the
following.
The probability that the random variable X takes on a value between x and x + Δx is
P(x < X < x + Δx) = P(X ≤ x + Δx) – P(X < x)
F ( x + ∆x ) − F ( x )
∆x
is the probability per unit length in the interval. If we now let Δx tends to zero, we have
F ( x + x ) − F ( x)
Lim = F ' ( x ) = f ( x)
∆x − 0 ∆x
x
Hence, F(x) = ∫−∞
f ( y )dy where F(-∞) = 0
110
F(x) represents the probability per unit distance along the x acis and
d
F(x) = F(x)
dx
This shows that the p.d.f of X is obtained by differentiating the distribution function of X.
(i) f(x) ≥ 0
∞
(ii) ∫−∞
f ( x) dx = 1
a ∞
(iii) P(X < a) = ∫−∞
f ( x) dx = 1 - ∫a
f ( x) dx = 1 – P(X > a)
Figure (2) represents a graph of the function f(x). the area enclosed by the curve f(x)
between a and b is
b
∫a
f ( x) dx = P(a ≤ X ≤ b) = F(b) – F(a).
111
Theorem
Proof:
A ⊂ B since x1 ≤ x2
Thus
P(A) ≤ P(B)).
x
(ii) Lim F(X) = Lim
x−∞ x −∞ ∫
−∞
f ( y )dy = 0
x
Lim F(x) = Lim
x−∞ x −∞ ∫
−∞
f ( y )dy =1
Examples
e − λx , x ≥0
Exponential Distribution. Let f(x) =
0, x<0
f(x)
112
Figure 3. Density function of exponential distribution.
The random variable X having its probability density function given by (1) is said to have
exponential distribution with parameter λ.
1 − 12 x 2
Let f(x) = e , − ∞ < x < ∞.
2π
113
5.5 Let the random variable X have the probability density function f(x) – 1/2e-1/2x 0
≤ x ≤ ∞. Find (i) P(2 ≤ X ≤ 7) (ii) P(X ≥ 10).
Solution
7
−1 / 2 x − 2x 7
P(2 ≤ x ≤ 7) = ∫ 1 / 2e dx = − e Ι = e −1 − e − 7 / 2 = 0.338
2 2
And
∞ ∞ −5
P(X ≥10) = ∫ 1 / 2 e −1 / 2 x dx = − e −1 / 2 x Ι =e .
10 10
f(x) =
0 otherwise
(i) what is the value of a? and (ii) find P(X > 1).
Solution
∞
∫−∞
f ( x) dx = 1
implying that
∞ ∞
∫ ae −2 x dx = 1 = a ∫ e − 2 x dx = 1 / 2
0 0
==>a=2
114
hence
f(x) =
0 otherwise
∞
−2 x
(ii) P(X > 1) = ∫ 2e
1
dx = e-2 = 0.14
Let X be a continuous random variable with probability density function f(x). Suppose we divide the
range of X into small intervals, each of length Δx (by the definition of f(x)).
∑ xf ( x) ∆x
The limiting value of this as Δx 0
∫ xf ( x) dx.
This leads us to the following definition.
Definition 3
∞
E(x) = ∫−∞
xf ( x) dx
In general,
∞
E(Xr) = ∫ −∞
x r f ( x) dx
115
∞
E(Φ(X)) = ∫−∞
Φ ( x) f ( x) dx.
∫ (x − µ )
2 2 2
The variance σ = E[(X – μ) ] = f(x) dx = E(x2) – μ2
−∞
Where μ = E(x).
Examples
5.6 Suppose X is a continuous random variable with p.d.f f(x) = λe-λx, 0 < x < ∞.
Find (i) the mean and variance of X and (ii) the probability distribution function F(x), and the probability
that X lies between 2 and 5.
Solution
∞ ∞
E(X) = ∫ x.λe −λx dx = ∫ xλ e −λx dx
−∞ 0
By integrating by parts
∞ d − λx ∞ ∞ 1
= ∫ −x (e ) = − xe −λx Ι + ∫ e −λx dx = .
−∞ dx 0 0 λ
Similarly,
∞ ∞ d − λx ∞ ∞
E(x2) = ∫ x 2 .λe − λx dx = ∫ − x 2 (e ) = − x 2 e − λx Ι + 2 ∫ xe −λx dx
0 0 dx 0
0 2 2
= 2∫ x o − 2 x dx = xλl − 2 x dx = 2 E ( x) =
o λ∫ λ2
116
2 1 1
hence, var(x) = − =
λ 2
λ 2
λ2
x
xf ( y ) dy = x − λy [ =1− e1 x
f ( y )dy = ∫ λl − λydy = −e
F(x) = ∫− ∫ 0
=
(u)
5
p (2 < x < 5) = ∫ λe 2 x dx = F (5) − F (2)1 − e −5 − 1 + e − µ = e − µ − e − µ
2
1
6
0<x<6
F(x) =
elesewhere
solution
1x
(1) F(x) = ∫0 6
= p( x≤ x)
0 1 1 1 2 6 1 1
E(x) = ∫0 x 6dx = 6 − 2 x [ 0 = 6 − 2 − 36 = 3
6 1 1 1 1 1
E ( x 2 ) = ∫ x 2 − dx = − x 3 [ 60 = − − 6 3 = 12
0 6 6 3 6 3
117
hence var(X) = 12-9 =3
definition 4
1 x2
1
for example. If f(x) = 2 xe isasynmericfunction.
2
similarly,
1
f(x) = −0< x < a
2x
Definition.
1
−1
F(x) = 2π e ( x − π ) 2 =< x <=
2
118
Then
1 1 ( µ + x- µ )2 1 − 1
F(( µ − x) = e = e
2x 2 2π 2
1 1 ( µ + x- µ )2 1 − 1
F( µ − x) = e = e
2x 2 2π 2
theorem 1
f(x) = 1-f(-x)
proof
x −x
f(-x)= ∫
−
f ( y )dy = ∫
−
f (− y )dy
119
x = = x
f(-x) = ∫
−
f ( z )(−1)dz = ∫ f ( z )dz =
x ∫
−
f ( z )dz − ∫ f ( z )dz = 1 − f ( x)
−
and hence
f(x) = 1-f(x)
terminology
one of the most important continuous random variables is the Normal random
variable, X. The range of X is the real line.
Definition
The Normal Distribution. The continuous random variable X has the normal
distribution with mean µ and varianceo3 if µ has the probability density function.
− 1 (x − µ )
3
1
F(x) = e =< x <=
2π 2 02
120
f ( x ) dx = 1
1. ∫
2. ∫ xf ( x) dx =µ
( x − µ ) 2 f ( x)dx = 0 2
3. ∫
Properties of the Normal Distribution.
f ( x ) dx = 1
1. ∫ this can be verified.
x−µ
Thus let z =
σ
∞ 1 −
1
(x − µ )2 dx = ∞ 1 −
1
∫
−∞
σ 2π
e 2
σ2 ∫
−∞
2x
e 2
dz = 1
1 1
∞ − y2 ∞ ∞ − (z2 + y2 )
1 1
I = ∫ dy = ∫ ∫
2 2 2
e e dz dy.
2π −∞ 2π −∞ −∞
Z = rcos θ, y = rsin θ
r2
1 2π ∞ 1 2π − ∞ 1 2π
I = 2
∫ ∫ re −r 2
dr dθ = ∫ −e 2
Ι dθ = ∫ dθ = 1.
2 0 0 2π 0 0 2π 0
(x − µ)2
1
∞ 1 −
∫
−∞
σ 2π
e 2
σ2
dx =1.
121
2. E(X) = μ, Var(X) = σ2.
1 − 12 x 2
f(x) = e ,− ∞< x<∞
2π
Most available tables are tabulated only for positive values of x since F(-x) can be
found from f(x) and f(x) are given at the back of this book. The graph of f(x) is
bell-shaped curve centred at its mean.
X −µ
4. IF X has a normal distribution with mean μ and variance σ2, then Y = has a
ο
standardized normal distribution. The proof of this is left as an exercise. Y is said
to have a standard Normal distribution.
(x − µ)2
1
∞ 1 −
Mx(t) = E(etx) = ∫ −∞
e tx σ
2π
e 2
σ2
dx.
X −µ 1
Let y = , then dy = dx.
σ σ
x = μ + y σ.
Thus,
1
∞ 1 − y2
Mx(t) = ∫ e t ( µ + yσ ) e 2
.σ dx dy.
−∞ σ
2π
122
e tµ
1 1 1
∞ − ( y − tσ ) 2 tµ + σ 2 t 2 1 ∞ − ( y − tσ ) 2
=
2x
∫ −∞
e 2
dy = e 2
2π
∫−∞
e 2
dy .
Let y – tσ = y, then
1 1
1 ∞ − ( y − tσ ) 2 1 ∞ − y2
2π
∫
−∞
e 2
dy =
2π
∫ −∞
e 2
dy = 1.
Hence,
1
tµ + σ 2 t 2
2
Mx(t) = e
Examples
5.8 If X has a standard normal distribution i.e. mean of X, μ=0, and variance = 1. Find
Solution
∞
(i) P(X ≤ 0.5) = ∫
−∞
f ( x) dx = F (0.5).
From table (3) F(0.5) is found by locating the first two digits (0.5) in the column
headed x, the desired probability of 0.6915 is found in the row labeled 0.5 and
the labeled 0.00. That is
1
0.5 1 − x2
P(X ≤ 0.5) = ∫−∞
2π
e 2
dx = 0.6915.
and
Example 5.9
Solution
L M / N HO L M / 0.5987
IJ K <J4 3
5 5 5
(i) P (X <(6) = PH
U U M / W'0.5 U O U 0.5X
>J 4 IJ4 VJ4
5 5 5
(ii) P(3 < X < 7) = PH
Notation:
When we write X is N (o2), we mean that X has a Normal probability distribution with
mean and variance o2
124
Theorem (De Moivre laplace Limit Thoerem)
If X denote a binomial random variable with mean up and variance np(1-p), then for any
a < b,
^ ' )_
lim N \] L L ab / ФWaX ' ФW] X.
0Z[ `)_W1 ' _X
The accuracy of this approximation is quite good for values of n greater than 10 P(1-p).
Example 5.10
Let X be a binomial random variable with parameters n = 40, p = 0.20. find the
probability that X = 5.
Since X is discrete and normal distribution is for continuous random variable, we write
U Y M
5.4J5 IJ5 4.4J5
[Link] [Link] [Link]
P (X = 5) = {(4.5 < 5.5 = PH
f W( X / g hi lf ( m 0 n
Jjk
0 (U 0
The pdf has the graph shown in Figure 3. It can easily be shown that
125
[
o hi Jjk / 1.
e
The exponential distribution plays an important role in point processes and in describing
random phenomena. A number of real life situations can be modeled using an
exponential distribution.
(1) E(X) and standard deviation of X are the same. The expected value of X is given by
Hence,
' / .
: 3 3
j, j, j,
Var(X) =
Thus, the mean and standard deviation of an exponential distribution are the same.
P (x > s) = 1 – F(x + t)
1 – F (t)
That is,
1 – F(s) = F(s+t)
1-F(t)
(1-F(s)) (1-F(t) = 1 –F (s + t)
1 – F (x) = ex
(p(x>t) = 1 – F (t) = e
P(X > s) = e
e-µ = e
Definition 6
I (x) =
(1) = pe i Jw qv / 1.
[
Hence
Γ(x) = (x-1)!
ΓH M / pe v J, i Jw qv / √{
3 [ -
:
(i)
pe / 1.
[ w |}- ~ }w
W|X
(ii)
Definition 7
Let X be a continuous random variable assuming only non-negative values. Then X is said
to have a gamma distribution if its pdf is given by
i Jjk ( Y 0
jWjkX|}-
f W( X / \ W0X n
0 ii
ii
λ and n are called the paramenters of the distribution. f(x) is denoted by Γ(n, λ).
E(X) = , ]W^ X /
0 0
j j,
1.
Hence,
128
. / .
j| W03X 0
j|- j
E(X) =
W0X
Similarly,
[
yW) = 2X
o ( 03 i Jjk q( / .
e h0:
Hence,
. / .
j| W0:X 0W03X
W0X j|, j,
E(X2) =
Thus
' / .
0W0 3X 0, 0
j, j, j,
Var(X) =
P(X ≤ x) = P(Y ≥ n)
That is
pe qv / ∑[ .
[ j| w|}- ~ } WjkX ~ }
W0X 0 !
3. If n = 1, f(x) becomes
129
f(x) = λe-λx.
( H M i J,k
| 3 , |
J3
,
1
/ ( , J 3 i J, , ( Y 0
|
:
f W( X /
yH M 2, y H M
0 |
0
: :
A continuous random variable having pdf f(x) given by (7) is said to have a chi-square
distribution with n degrees of freedom (denoted by χn). Because of its importance, the
chi-square distribution is tabulated for various value of the parameter n) see table 4). If
X has a chi-square distribution with n degrees of freedom, we have
(i) E(X) = n;
130
5.2.7 The uniform distribution
A continuous random variable X assuming all values in the interval (a, b) where both a
and b are finite is said to be uniformly distributed over the interval (a, b) if its pdf is
given by
] L ( L an
3
f W( X / J
0 ii
ii
The graph of f(x) is shown in figure below:
131
, k WX /
~ J ~
: W J X
1. E(X)
i / 1 = ] = =
WX,
:!
i = 1 = a = =
WX,
:!
I WX / x = =
3 r , J , s W A J A X ,
J :! >!
/ /
W , J , X WJXW X
:WJX :W J X :
Mx’(0) =
Similarly,
W A J A X
>WJX
Mx’(0) =
Hence,
] L ( L an
kJ
F(x) = P(X ≤ x) = \ J
1 (Ya
A continuous random variable X is said to have a beta distribution if its pdf is given by α,
β > -1
( W1 ' (X , 0 U ( U 1n
3X!
f(X) = \ ! !
0 ii
ii
f(x) becomes the pdf of a uniform distribution over (0, 1) when α = β = 0.
/ pe ( W1 ' (X q(
! ! 3
W 3X!
and
3 W = ¢ X! £!
o ( ¡ W1 – ¢X q( /
e W = ¢ = £ = 1X!
We have
9 / .
W 3X! W¡ X! ! W 3X!W¡ X!
! ! W¡ 3X! !W¡ 3X!
E(xr) =
A random variable is said to have a Cauchy distribution with parameter θ, -∞ < θ < ∞, if
its pdf is given by
f(x) = . , '∞ U ( U ∞
3 3
¤ 3 WkJ ¥X,
Exercises
133
§( 0 U ( U 1n
f W( X / ¦
0 ii
ii
Determine the value of the constant c. find F(x) and the mean and variance of X.
( 0 L ( L 1
f W(X / ¨2 ' ( 1 ( L 2n
0 ii
ii
5. Let X be a Normal random variable with mean 8 and variance 2. Find P(-2<X<6).
Determine the values of a and b such that P(X< a) = 0.25 and P(X>b) = 0.25
6. Suppose that X is N(5, 9). Find a constant c such that P(X > c) = 2P(X< c).
7. Let a point be chosen uniformly in the interval (a, b). ). Let let X denote the
distance of the point chosen from a. find the probability distribution function F(x)
of X.
10. The lifetime in hours of an electric light bulb is a random variable with probability
density function
134
0, (U0
¯(, 0U(U1
f W( X / n
® :, 1U(U§
3
0, ( Y§
Where c is a constant
(i) Calculated the value of c and hence the mean and variance of X.
(ii) Sketch the graph of f.
12. A random variable X has pdf.
1 – F(x)
15. (a) If X is N (o2) find k (as a function of u and 02 such that P(X > k) = 2P (X <k)
Find the M.G.F of X and hence the mean and variance of x in each case
17. A random variable X has mean 5 and variance 3. If the pdf of X is given by f(x) =
3
J
, if a < x < b and zero otherwise. Find a and b
135
18. If f(x) = in 1/x, 0 < x < 1 and 0, otherwise show that f(x) is a pdf
¢ ik ' ∞ U ( U £ n
19. Let X be a random variable with zero expectation and probability density function
given by f(x) = ¦
0 ii
ii
Find α.
20. Let f(x) = ½ (k) = ½ (k-1)/ (1+/x/)k, k ≥ 1-∞ < x< ∞show that f is pdf
136
UNIT 2
JOINTLY DISTRIBUTED RANDOM VARIABLES
6.1 Bivariate Distribution
In our study so far, we have considered only univariate distribution, Univariate
distribution depends on only one random variable but we are frequently interested in
specifying two or more random variables for a member of a population and interested in
probability statements. For example, the age A and the systolic blood pressure, P may be
of interest and we would consider (A,P) as a single experimental outcome. We might
study the height H and the weight W of a chosen person given rise to the outcome (h, w).
A probability distribution which depends on two random variables. Table (1) gives the
joint variables. Table (1) gives the joint probability density function of two discrete
random variables X, number of white marbles, and Y, number of blue marbles in a
sample of 3 chosen at random from a box containing 4 red, 3 white and 5 blue marbles.
Table 1
X/Y 0 1 2 3
3 1/220 0 0 0 1/220
We denote the probability that X takes the value of x and Y the value y by f(x, y) where
f(x, y) = P(X = x, Y = y).
For example, the probability that the sample contains 1 white and 2 blue marbles is
denoted by f(1, 2). From the table,
137
f(1, 2) = 3/22
This is calculated thus:
P(X = 1, Y = 2) =
Table 2
x 0 1 2 3
y 0 1 2 3
If we know the joint probability density function of two random variables X and Y, then
we can compute the marginal probability density function fx(x) by summing over y and
the marginal probability density function fy(y) by summing over x i.e.
fx(x) =
fy(y) = (1)
Note: It may happen that X is discrete while Y is continuous, but in most applications it
is either both are discrete or both are continuous.
138
Definition 1
Two discrete random variables X and Y are said to be independent if their joint density
function is give by
f(x, y) = fx(x) fY(y)
i.e.
P(X = x, Y = y) = P(X = x) P(Y = y) for all pairs of x, y.
fy(2) =
f(1, 2) =
and
E(X + Y) = E(X) +E(Y) = 2.99.
Expectation of Product
The expectation of product of two discrete random variables is defined as
139
E(XY) =
+ + +
+ + ×
= +
Definition 2: Covariance
The covariance of two random variables X and YU denoted by Cov(X< Y) is denoted by
Cov(X, y) = E{(X –E(X))((Y – E(Y)}
= E[XY – XE(Y) – YE(X) + E(X)E(Y) = E(XY) – E(X) E(Y)
Example 1
Given the following joint probability density function of X and Y. Find
(i) The marginal probability density function of X and Y
(ii) The expectation of X and Y
(iii) The covariance of X and Y
140
(iv) The variance of X + Y.
X/Y 1 2 3
3 .1 .2 0
4 0.2 0 .1
6 .15 .15 .1
fX(x) .3 .3 .4
Example 2
Two random variables X and Y have the joint probability distribution given by
X/Y 1 2
1 0.2 0.4
2 0.3 0.1
And that of Y is
Y 1 2
142
P(X = 1, Y = 1) = 0.2
P(X = 1) = 0.6, P(Y = 1) = 0.5
P(X = 1) P(Y =1) = 0.30 ≠ P(X = 1, Y =1) = 0.2
Hence X and Y are not independent.
Consider two independent random variables X and Y. then
E(XY) = E(X)E(Y).
To see this, note that
E(XY) =
ρ = ρ (X, Y) =
Definition 4
If X and Y are two jointly distributed discrete random variables, trhe conditional
probability density function of X given Y = y is denoted and defined by
fXDy(xDy) = P(X = xDY = y) =
It follows that
fXDy(xDy) =
And
F(x, y) = fXDy(xDy)fy(y).
Example 3
Let the joint probability density function of X and Y be given by the table below.
X/Y 1 2 3
3 .1 .2 0
4 .2 0 .1
6 .15 .15 .1
fXDy=2(3D2) =
fXDy=2(4D2) =
144
fXDy=2(6D2) =
fXDy=2(xD2) 0
It is left as an exercise to show that the conditional pdf of X given that Y= 1,3 are
1. fXD1(xD1)
x 3 4 6
2 4 3
9
fXD1(xD1)
9 9
2. fXD3(xD3)
X 3 4 6
fXD3(xD3) 0
Example 4
Calculate the conditional expectation and conditional variance of X given that Y = 2
where the joint pdf of X and Y is given in Example 3. From example 3, we have
145
E(XDY = 2) =
E(X2DY = 2) =
Exercise 6.1
1. Two discrete variables X and Y have the joint probability distribution given
below.
X/Y -1 5 7
0 0.1 0.2 K
1 0.05 0 0.15
2 0.3 0.1 0
(a) Find the value of k, (b) Find the marginal distribution of X and Y and
2. Given the following joint probability density function of random variables X and
Y.
X/Y 1 2 3
1 1/8 1/8 ¼
2 3/8 0 0
3 0 1/8 0
Find
146
1. The marginal probability density function of X and Y and hence E(X) and
E(Y).
2. Are X, Y independent?
3. Find the conditional pdf of Y given that X = x for which E(YDx) is defined.
X/Y -1 0 3
2 .1 0 .1
3 .2 .1 0
4 0 .3 .2
4. Suppose a box has 3 balls labeled 1, 2, 3. Two balls are drawn at random one after
the other without replacement. Let X and Y denote the number on the first and
second balls drawn respectively
5. Suppose the situation is as an exercise 4, except now that the two balls are selected
with replacement.
X/Y -3 -1 0 2 5
148
Where c x1, x2, … xk) = number of possible sequences of outcomes for the n experiments
to yield
X1 = x1, X2 = x2 = … Xn = xn.
From our knowledge of mathematics of counting,
c x1, x2, … xk) =
can be regarded as the number of ways n objects can be partitioned into k classes such
that class contains xi objects (I = 1, 2, …k). Thus,
P(X1 = x1, X2 = x2,… Xk = xk) =
Example 5
Suppose that a fair die is rolled 12 times. Find the probability that 1 and 6 appear 2 times
each. 3 thrice each, 3 times, and 4 and 5 once each.
Solution
Let Xi be the number of times ith outcome appears.
n=12, x1 = 2, x2 = 3, x3 = 3, x4 = 1, x5 = 1, x6 = 2, pi =
Hence,
P(X1 = 2, X2 = 3, x3 = 3, x4 = 1,x5 = 1, x6 = 2)
Example 6
In a certain large population, 70% are right-handed, 20% left handed and 10% are
ambidextrous. If 10 persons are chosen at random from the population, what is the
probability that (i) all are right-handed? (ii) 7 are right-handed, 2 are left handed and 1 is
ambidextrous.
Solution
149
P1 = 0.7,P2 = 0.2,P3 = 0.1
Let X1, X2, X3 denote number that are right-handed and ambidextrous respectively, then
and
(that is, summation in discrete case is replaced by integrals) and the marginal probability
density functions are given by
fX(x) =
fY(y) =
F(x, y) = P(X ≤ x, Y ≤ y)
F(x, y) =
150
If B is in the range space of (X, Y)
and
From the joint probability distribution function we can get the marginal distribution
functions by the following relationship
FX(x) = F(x, ∞) =
FY(y) = F(∞, y) =
Let X and Y be any continuous random variables. Then X and Y are independent if and
only if
E(YDX=x) =
151
Example 7
Solution
The marginal probability density function is given by
(i) fX(x) =
fX(x) =
and
fY(y) =
=
Hence
2(1 – x) 0<x<1
fx(x) =
0 elsewhere
and
2y 0<y <1
fy(y) =
0 elsewhere
(iii) E(XDy) =
=
Example 8
Let the random variables X and Y have the joint pdf.
Solution
(i) fX(x) =
=
fY(y) =
=
Hence,
+x 0<x<1
fX(x) =
0 elsewhere
+y 0 <y< 1
fY(y) =
0 elsewhere
(iii) E(X) =
E(Y) =
E(Y) =
Hence
E(XY) = 1/3
153
Thus,
Cov(X, Y) = 1/3 – 7/12.7/12. = -1/144.
Example 9
Let the random variables X and Y have the joint pdf
0 <y< x < 1
f(x, y) =
0 elsewhere
Find the marginal p.d.f of X and of Y and the conditional pdf of Y given X = x.
Solution
Therefore
1, 0<x<1
fX(x) =
0 elsewhere
FY(y) =
Therefore
-log y 0 <y< 1
fy(y) =
0 elsewhere
0 <y< x < 1
fYDx(yDx) = =
0
Hence X and Y/x are uniformly distributed over (0, 1), (0, x) respectively.
Example 10
Let the random variables X and Y have the joint pdf
x2+ 0 ≤x ≤ 1, 0 ≤ y ≤ 2
f(x, y) =
0 elsewhere
Find the P(X + Y ≤ 2).
154
Solution
Let A ≡{x, y: x + y ≤ 2}.
P(A) =
=
Exercise 6.2
1. Suppose that X and Y have the joint p.d.f
K e-λ(x+y)x≥ 0, y ≥ 0
f(x, y) =
0 elsewhere
Find (i) K and (ii) the marginal pdf of X and of Y.
2. Suppose that X and Y have the joint pdf
C(x-y) 0 <x <2
f(x, y) = -x< y < x.
0 elsewhere
0 ≤x ≤2
f(x, y) =
0 elsewhere
(i) Find K using the fact that ∫ ∫ f(x, y) dxdy = 1.
(ii) Are X and Y independent random variables.
4. Two random variables X and Y have the joint pdf given by
c(x – 2xy + y)0 ≤x ≤ 1, 0 ≤ y ≤ 1
f(x, y) =
0 elsewhere
(a) Show that X and Y are uniformly distributing over [0, 1].
(b) Find the joint probability distribution function of X and Y. Hence, find the
marginal distribution functions of X and Y.
5. Suppose that X and Y have the joint pdf
λ2e-λy 0 ≤x ≤y
f(x, y) =
0 elsewhere
(i) Find the marginal probability density functions of X and of Y.
6. Suppose that X and Y have the joint pdf
11. If it is assumed that, in single – car accidents in a certain area, the probabilities of
minor severe, and fatal injuries to the driver are 0.5, 0.4 and 0.1 respectively, find
the probability that in 10 accidents there are 4 minor, 4 severe and 2 fatal injuries
to the 10 drivers.
156
Example 2
Let X have the poisson pdf with parameter λ. Find the pdf of Y = X2
; ^ / 0, 1, 2, …
~ } j±
k!
f(x) =
Solution
=f(`v), x = `v
(single-valued inverse function)
~ Jj j`
=¨ , v / 0, 1, 4, 9, 16 …n
√w
0 ii
ii
Since X is a positive random variable), y = 0, 1, 4, 9, 16,…
Example 3
Suppose X is a random variable having probability density function f(x) given by
X -2 -1 0 1 2 3
f(x) .15 .25 0.15 .2 .15 .1
Solution
Possible values of Y are 0, 1, 4, 9.
This is not a one-to-one function
P(Y = 0) = P(X = 0) = 0.15
P(Y = 1) = P(X = -1 or 1) = P(X = -1) + P(X = 1) = 0.25 + 0.2 = 0.45
P(Y = 4) = P(X = -2 or 2) = P(X = -2) + P(X = 2) = 0.15 + 0.15 = 0.3
P(Y = 9) = P(X = 3) = 0.1
Hence, the pdf of Y is given by
y 0 1 4 9
h(y) 0.15 0.45 0.3 0.1
Example 4
3 k
f W( X / H M , ( / 1,2
Let X has the pdf given by
:
0 elsewhere
Find the pdf of Y defined by
1 if X is even
Y=
1 if X is odd
Solution
157
P(Y = 1) = P(X = 2, 4, 6, 8,…
3 : 3 5 3 <
= P(X =2) + P(X = 4) + P(X = 6) +…
/ H M = H M = H M = …/
3
: : : >
P(Y = -1) = .
:
Similarly,
>
Hence, the pdf of Y is given by
1 2
Y 1 1
3 3
g(y)
Continuous Case
Let X be a continuous random variable having probability density function f(x). we wish
to find the probability density of a variable Y = g(X). We shall discuss methods for
finding the pdf of Y.
Let FY(y) be the probability distribution function of Y. Then, expressing FY(y) in terms of
X, we obtain
FY(y) = P(Y ≤ y) = P(g(X)) ≤ y)
Let X be uniformly distributed over (0, 1). Find the pdf of Y = log~ W1 ' (X, ( Y 0.
J3
Example 5
j
Hence
FY(y) = (1 - e-λy)
:
Hence,
i J,WwJ3X , v m 1
3 -
:
fY(y) =
0 elsewhere
Example 7
λ( jJ3 0≤x≤1
let X has the density function given by
fx(x) =
0 elsewhere
Show that Y = - loge X is an exponential random variable.
Solution
µ
P(X ≤ u) = pe h( jJ3 q( / ( j | / µ j
P(Y ≤ y) = P(-loge X ≤ y) = P(X ≥ e-y) = 1 – P(X ≤ e-y)
´
0
thus,
P(X < e-y) = e-λy
P(Y ≤ y) = 1 – e-λy; fY(y) = λe-λy, y ≥ 0
Hence,
FY(y) = P(Y ≤ y) = 1- e-λy,
λe-λy , y≥0
Fy(y) = F’Y(y) =
0 elsewhere
This shows that Y is an exponential random variable.
Example 8
159
Let X be a random variable having pdf
2x 0≤x≤1
fx(x) =
0 elsewhere
Find the pdf of Y = 4x2.
Solution
w
FY(y) = P(Y ≤ y) = P(4X ≤ y) = PH^ : ≤ M
2
5
w J w √w √w J√w
= PH^ :
≤ M / PH √ ≤ ^ ≤ M= F XH M - FXH M
5 : : : :
´ ´ µ
FX(u) = pe fW(Xq( / pe 2( q( / ( : |0 / µ :
Thus,
:
w J√w
FY(y) = FXH√ M ; FXH M /0
: :
Example 9
Let X be a continuous random variable having probability density function f(x). derive an
expression for the probability density function of the random variable Y = X2.
Solution
FY(y) = P(Y ≤ y) = P(X ≤ y) = W'`v ≤ ^ ≤ = `vX = FX(+`v) - FXW'`v).
2
Differentiation gives
3 J3
F’X(`v) = fXW`vX, F’XW'`v) = fXW'`vX.
:√w :√w
Hence
3
(fXr`vs + fXr'`vs; y > 0
:√w
F’Y(y) = fY(y) =
0 elsewhere
Example 10
Lex X be a random variable having the pdf.
3
(x + 1); -1 ≤ x < 1
:
f(x) =
160
0 elsewhere
2
Find the pdf of Y = X .
Solution
From example 10above, we have
3
(fXW`vX + fXW'`v)
:√w
fY (y) =
fX(`v) = W`v = 1X
3
:
W1X / \ : v 0 U v U 1 n
3 -
hence,
ii
ii
:√w : : :√w
0
Note: The transformation Y = X2 mapped x:-1 ≤ x < 1 onto the set y: 0 < y < 1.
Example 11
Let X be a continuous random variable having pdf f(x). Derive the formula for the pdf of
Y = |X|
FY(y) = P(Y ≤ y) = P(|X| ≤ y) = P(-y < X < y) = FX(y) – FX(-y)
By differentiation we see that
Theorem 1
Let Y = g(X) be a differentiable strictly increasing or strictly decreasing function on an
interval D and let D * denote the range of g. let X e a continuous random variable with
pdf f(x) such that f(x) = 0 for x ϵ D. Then the pdf of Y is given by
fY(y) = fX(g-1(y)) ¹ ¹ ; v » ¼*
º}- WwX
w
fY(y) = f(x) ¹ ¹
-1
k
where x = g (y) is the inverse function of g. equivalently
w
where x is expressed in terms of y, y ϵ D* and x = g -1(y).
Proof:
Case 1: suppose that g(X) is strictly increasing
FY(y) = P(Y ≤ y) = P(g(X) ≤ y) = P(X ≤ g -1(y)) = FX(g-1(y))
Thus by differentiation, we have
w
F’Y(y) = fX(g-1(y)) (g-1(y)).
Since g(X) is strictly increasing so also is g-1(y).
161
½J3 WvX¹ / fI W( X ¹ ¹ ; ( / ½J3 WvX.
k
FY(y) = fX(g -1(y))¹
w w
Example 12
Let X be a random variable having the pdf
1 0<x<1
f(x) =
0 elsewhere
Find the pdf of Y = -2 loge x.
Solution
w : w :
3
Hence
:
fY(y) = 1, e-y/2, y ≥ 0.
3 -y/2
Thus
:
e y>0
fY(y) =
0 elsewhere
Example 13
Let X be a random variable with pdf
2x 0≤x≤1
fX(x) =
0 elsewhere
2
Find the pdf of (i) Y = 2X – 5 (ii) Y = X .
Solution
(i) Y is a strictly increasing function of X. thus,
162
f¿ WvX / fI W( X ,( /
k w4
w :
/
k 3
w :
f¿ WvX / 2(, / ( /
3 w4
Thus,
: :
' 5 ≤ v ≤ '3
w4
Hence
:
fY(y) =
0 elsewhere
q( 1 -
(ii) Y is strictly increasing on (0, 1) and 0 ≤ Y ≤ 1.
^ / = `v ])q / v J,
qv 2
f¿ WvX / fI W( X
k
thus,
w
f¿ WvX / 2( ¹ v ¹ / 2. v . v J, , ( / `v / 1.
3 - - 3 -
J J
, ,
: :
Hence,
1 0≤y≤1
fY(y) =
0 elsewhere
Let X be a random variable having the N(µ, σ2). Find the pdf of (i) O /
IJ À
Example 14:Chi-square Distribution
Á
(ii) Z = Y2.
O/ / Â
IJ À k
Solution
Á w
(i) , X = µ + Yσ,
fY(y) = fX(x) ¹ ¹; x µ + yσ
k
Y is a strictly increasing function of X, hence
w
fI W( X / i J, .
3 - Wk J ÀX,
Á √:¤ Á,
fI Wμ = vÂX / i J,w
- ,
Substituting for x, we see that
3
Á √:¤
Á √:¤ Á√:¤
That is Y is a standard normal random variable (N(0, 1)).
(ii) Z = Y2
163
= f¿ WvX ¹ ¹
0 w w 0
fZ(z) = fY(y) ¹ ¹
Ä '∞ Ä '∞
v / Å,, / Å ,
- w 3 -
Ä :
√:¤ : √:¤ :
- - -
0U Å U ∞n
- - } É
H M, È , ~ ,
i / Ç
} -
Ä , J Ä ,
,
√:¤ √¤
ii
ii
=
0
Definition 3
( , J 3 i J, n , 0 U ( U ∞
-
H M,
fW( X / Ç
|
,
|
H M
,
0 elsewhere
is said to have a chi square distribution with n degrees of freedom and is written ^ : W0X .
IJµ :
Then the pdf of O / H M l^ : W3X.
Á
Theorem 2
Let X be a continuous random variable having pdf f(x) such that f(x) > 0, x » D and f(x) = 0, x ∉
D. Suppose D can be partitioned into two sets D1 and D2 such that the transformation Y = g(x)
which is a monotone differentiable function in D and is a strictly increasing function on D1 and
strictly decreasing differentiable function of X on D2, then the pdf of Y is given by
164
Consider the following examples.
Example 15
i J,k ' ∞ U ( U ∞
- ,
3
fW( X /
√:¤
165
(i) The graph of Y is given below.
Figure 1
/ 1.
k
w
On D1, y = x,
fI WvX ¹ ¹ / i J,w
- ,
k 3
w √:¤
/ 1.
k
w
On D2, y = x,
fI W'vX ¹ ¹ / i J,w .
k 3 - ,
w √:¤
Thus,
w Ì w Ì √:¤ √:¤
- ,
Hence
f¿ WvX / fI W( X ¹ ¹ = fI W( X ¹ ¹
k k
w Ë0Ì w Ë0Ì
- ,
166
1
i J ,k
- ,
f W(X /
√2{
( / E`v
√ :¤
On D2, x = + `v])q
fI r`vs ¹ ¹ / i J,w . v J,
k 3 -
3 -
w √:¤ :
Hence
- -
vm0
}
w,~ ,
0 ii ii
Example 16
Let X be a uniform random variable over the interval H , M. Find the probability density
J¤ ¤
: :
function of Y = tan X.
Solution
1 '{ {
fk W^ X / ¨{ 2 U ( U 2n
0 elsewhere
.
3 3
¤ 3 w,
167
3
f¿ WvX /
¤W3 w , X
-∞< y < ∞
Exercise 7.1
, ( / 1, 2, … ÒWiÓi)Xn
:
fW( X / \:Ñ-J :
0 elsewhere
2 if X is even
Y=
5 is X is odd
fW( X / g2(i
Jk ,,
0 U ( U ∞n
0 elsewhere
4. Suppose that X is a continuous random variable defined over (0, 1) such that P(X ≤ 0.45)
= 0. Let Y = 1 – X. Find the value of K such that P(Y ≤ k) = 0.40.
Suppose that X has the pdf f(x) = e-x, x > 0. Find the pdf of the following random
(b) Å / .
3
5.
WI3X,
variables, (a) Y = X3
6. Suppose that X is a geometric random variable with parameter p. Find the pdf of
168
(a) Y = X2 (b) Z = X + r, where r is a constant.
10. Let X be uniformly distributed over (0, 1). Find the pdf of
(a) Y = eX (b) Z = X2 + 1
3
(c) Ð/ I 3
(d) Y = X1/β, where β ≠ 0.
11. Let X be exponentially distributed with parameter λ. Find the pdf of Y = loge X.
12. Let X be have gamma density with parameters (α, λ). Find the pdf of
15. Let X and Y be independent random variable each having uniform pdf on (0, 1,…N).
Find the pdf of (i) min (x, Y) (ii) Max (X, Y).
The method discussed in section 7.1 of finding the pdf of a function of one random variable of
the continuous type will now be extended to functions of two random variables
The problem of finding the pdf of Z is somewhat more involved. The following steps are useful.
169
2. Obtain the joint pdf of Z and W, say d(z, w).
3. Obtain the desired pdf of Z by simply integrating d(z, w) with respect to w. that is
∫ d(z, w) dw.
(i) The functions h(X, Y) and k(X, Y) satisfy the following conditions.
may be uniquely solved for x and y in termsof z and w, ssay x = g1(z, w),
y = g2(x, w)
Ôk Ôk Ôw Ôw
, , ,
ÔÄ ÔÕ ÔÄ ÔÕ
(b) The partial derivatives exist and are continuous
From the result of section 7.1 it can be proved using results of advanced calculus that under the
assumptions (i) and (ii), the joint pdf of Z and W is given by d(z, w) = f(x, y) |Ö| where x = g1(z,
w), y = g2(z, w) and
Ôk Ôk
Ö / ×Ôw
ÔÄ ÔÕ
Ôw
×.
ÔÄ ÔÕ
The determinant J, is called the Jacobian of the transformation (x, y)- (z, w). some of the
important random variables we shall be interested in are
X +Y, XY, X/Y, min. (X, Y), max. (X, Y), etc.
Distribution of Sum
0 1
Ö / ¹ ¹ / '1
1 '1
Thus d(z, w) = f(x, y) |Ö| = f(w, (z – w)),
170
Example 17
1 0<x<1 0<y<1
Solution
0 1
Ö / ¹ ¹ / '1
1 '1
Hence, d(z, w) = 1.1 = 1.
f(x)
x
Figure 2
171
Thus, the boundaries of B are the lines w = 0, w = 1, z = w, z = w + 1. This is show in the figure
below.
z
z =w+1
w=1
B
z =w
w
Figure 3
Hence
d(z, w) = 1
thus, (z, w) ∈ B
qWÅX / pe 1 q / Å
Ä
0<z<1
Z 0<z≤1
0 elsewhere
Example 18
Suppose a point (x, y)is selected at random in a circle centre (0, 0) and radius 1 and that this
point is uniformly distributed over the circle. Find the pdf of
d(z)=
0 elsewhere
172
Let W = tan-1(Y/X) then tan W = Y/X, x = r cos w, y = r sin w.
y
r
w 1
x
B
2 x
Figure 4
The Jacobian is
Ú( Ú(
Ö / Ù Ú Ú Ù ¹cos ' sin
¹ / W§Û : = l) : X / .
Úv Úv sin cos
Ú( Ú
hence
(r, w) » B
3
¤
r
Thus,
2 0LL1n
d(r) = ∫ d(r, w) dw = r pe q / ¦
:¤ 3
¤ 0 ii
ii
Example 19
3x 0<y<x
0 elsewhere
Let W = X, A = {(x, y), 0< y < x, 0 < x < 1} and B ={(z, w) : z=x – y, w = x}
173
y Z
A B
x W
1 1
Figure 5
The Jacobian is
1
Ö / ¹ ¹
0
'1 1
hence
3w 0<z<w
0 elsewhere
qWÅX / / pÄ 3 q
3
W1 ' Å : X
>
:
0<z<1
d(z) = 0 elsewhere
Distribution of Product
Suppose X and Y are continuous random variables having the joint pdf (x, y). show that the pdf
of Z = XY is given by
p ¹k¹ f H(, kM q(
3 Ä
Solution
Ä
Let W = X, thus, x = w, y = Õ. The Jacobian
174
1 0
Ö / Ü JÄ
3
3Ü / Õ
Õ, Õ
Hence,
Thus,
qWÅX / p ¹ ¹ f H, M q
3 Ä
Õ Õ
(2)
Example 20
Solution
Thus, from 2.
Ä >
qWÅX / p ¹Õ¹ . 2 34 HÕM q / Å > p ÕA q.
3 5 ; 3
34
Let A = {(x. y): 0< x < 1, 1 ≤ y ≤ 2}. The transformation, Z = XY, W = X gives the set B
determined as follows:
When x = 0, w = 0; when x = 1, w = 1,
175
The lines w = 0; w = 1; x = w, z = 2w are the boundaries of B as shown in the figure below.
B
1
Figure 6
Thus
pÄ/: ÕA q / Å, 0UÅU1
;Ä A Ä 3 5
qWÅX /
34 4
0UÅU1
5Ä
qWÅX / Ç 5Ä
4 n
W4 ' Å :X
1UÅU2
34
Distribution of Quotients
Suppose X and Y are continuous random variables having the joint pdf f(x, y). show that the pdf
of Z = Y/X is given by
J∞
1 0
Ö / ¹ ¹ / .
Å
Hence
Thus
176
Example 21
Let X and Y be independent random variables having the respective gamma densities Γ(α1, λ)
and Γ(α2, λ). Find the probability density function of Z = Y/X.
jÝ- k Ý- }- ~ }
fI W(X /
W- X
x>0
f¿ WvX /
jÝ, k Ý, }- ~ }
W, X
y>0
p ( i q
jÝ- Ý, È Ý,}- [ - Ý J3 JjkW3 ÄX
,
W- X W, X e
=
But
pe - , J3 i JjÕW3 ÄX q /
[ W- , X
WjW3ÄXXÝ- Ý,
Hence
,
W- , X È Ý, }-
W- X W, X W3ÄXÝ- Ý,
fZ(z) = z>0
Example 22
Let X and Y be independent exponential random variables with parameter λ. Find the
I
I ¿
joint pdf of Z = X + Y and W = .
Solution
The joint pdf of X and Y is given by
fX,Y (x, y) = λ2 e-λ(x + y) x > 0, y > 0
The transformation gives
x = wz, y = z(1 – w)
|Ö| = z
The Jacobian of the transformation is
Åh: i JjÄ , Å Y 0, 0 U U 1n
fÈ,Õ WÅ, X / g
The joint pdf of Z and W is given by
0 , ii
ii
177
Example 23
3
x ≥ 1, y ≥ 1, find the joint pdf of Z = XY
k, w ,
Let X and Y have the joint pdf fX,Y(x,y) =
and Ð /
I
¿
.
Solution
v (
From the transformation z = wy, w = x/y, we have
Ö / Ü 3 JkÜ /
J:k
w
w w,
|Ö| / 2.
Let X1 and X2be random variable with joint pdf fk-, , , (x1, x2). Let Y1 = X1 + X2,
Example 24
'
, Ôk,
Ôw- Ôw, : :
: 1, (2 2
Example 25
Let X and Y be independent exponential random variables with parameters λ1 and λ2
I
respectively.
I ¿
(i) find the joint pdf of U = X + Y and V =
(ii) show that U and V are independent if λ1 = λ2.
Solution
x = uv and y = u – uv.
178
f´ Wµ X / pe µh: i Jj´ qÓ / µh: i j´ , µ m 0
3
fß WÓ X / pe µh: i Jj´ qµ / 1 0 U Ó U 1.
[
Example 26
f W(, v X / gi ^ Y 0, v Y 0n
JWkwX
The joint pdf of X and Y is given by
0 Û
ili
find (i) P(X > 2, Y > 3), (ii) P(Y < X) and (iii) the pdf of Z = X/Y.
Solution
(: = v : ≤ : n
3
Example27
Let f(x, y) = \ ¤¡ ,
0 ii
ii
find (i)fX(x) and (ii) fZ(z) where Z = √^ : = O : .
√ : = ( : / :
√ : ' ( : n
Solution
fk W( X / pJ[ f W(, vXqv / p , qv / â
[ 3 w
¤¡ ¤¡ ,
√ : ' ( : ¤k ,
(i)
= pe pe i q(qv / 1 '
[ wÄ JWkwX 3
Ä3
fÄ WÅX / , 0 U Å U ∞.
3
Differentiating we obtain
WÄ3X,
Example 28
FZ(z) = P(X + Y ≤ z) = àkwäÄ f W(, vXq( qv / pJ[ pJ[ f W(, vXq( qv.
[ ÄJw
Distribution of sum. Let X and Y have the joint pdf f(x, y) and let Z = X + Y then
√¿
where X and Y are independent.
179
and W = Y then y = w and x = Å .
I √ß √Õ
Solution
Let Æ /
√¿ √ß
1 0
Then Jacobian of the transformation is
Ö / å' Å3 √ Õ √ Õå / √w
.
:
A √Õ
}
æ , √ß
√Ð 1 √Ð
è}-
The marginal pdf of z is given by
-É, ê ê [ É,
qÄ WÅX / o o
ê
i . q / i q
,
J J
è}- J ë3 ì
, è
√ √
, è , ,
√2{y H : M 2 √2{y H : M 2
æ | æ è
, , e
H1 = M; / H1 = M, /
Õ Ä, ´ 3 Ä, :´
: ß Õ : ß í,
3
Let u =
ç
2 2µ qµ
è}-
[
q È WÅ X / o î ï i J´ 3
,
√2{Óy H : M 2 1= H1 = M
æ è
J3 Ä, Ä,
, e
æ : æ
1
è}-
J3
oµ i J´ qµ.
,
è}-
√2{y H : M 2, H1 = M
æ è Ä, ,
æ
pe µ, i J´ qµ / y H M.
[ è ß3
Note
J3
:
Hence
è-
qÈ W Å X /
HM
. , '∞ U Å U ∞.
, 3
ç è-
√:¤ßH , M É, ,
ë3 ì
ç
The above distribution, dZ(z) is called a t-distribution.
Note
(i) That a t-distribution is completely determined by the parameter v, the number of
degrees of freedom of the random variable that has the chi-square distribution.
(ii) The t-distribution is a bell-shaped distribution, symmetric about zero.
(iii) The t-distribution is flatter than the Normal distribution. As the sample size
becomes large, t-distribution approaches the standard normal distribution (n > 30
is sufficient).
(iv) The values of the t-distribution for various values of degrees of freedom have been
tabulated.
180
Example 30
Exercise 7.2
1. Let X1 and X2 be independent normal random variables N(µ 1, σ21), N(µ 2, σ22)
respectively. Find the pdf of Y = X1 + X2.
I
2. Let X and Y be independent random variables having the respective pdfs.
f(x) = e-x, x ≥ 0, g(y) = 2e-y y ≥ 0. Find the pdf ofÆ /
¿
3. Let X and Y have the joint pdf f(x, y) = λ2 e-λ(x+y), x > 0, y > 0. Find the pdf of
Z = X + Y.
4. Let X and Y be independent random variables each having the normal probability
density function N(0, σ2). Find the probability density function of
¿ ¿, ¿ |¿|
I I, |I| |I|
(i) (ii) (iii) (iv)
5. Suppose X and Y are independent normal random variables with mean 0 and
variance σ2. Find the pdf of (i) X + Y (ii) X2 + Y2 .
Æ / |O ' ^ |.
7. Let X and Y have uniform distribution over the interval (a, b). Find the pdf of
8. Let X be uniformly distributed over (0, 1) and let Y be uniformly distributed over
(0, x). Find
(i) The conditional pdf of Y given X = x.
(ii) The joint pdf of X and Y
(iii) The marginal pdf of Y.
9. Let f(x, y) = c(y – x)2, 0 ≤ x < y ≤ 1
= 0, elsewhere
Find the pdf of Z = X + Y.
Theorem 3
Suppose that a random variable X has mgf Mx(t). Let Y = aX + b. Then, the mgf of the
random variable Y is given by MY(t) = ebt Mx(at).
Proof:
MY(t) = E(ety) = E(e(aX + b)t) = E(eatX + bt) = E(eatX ebt) = ebt E(eatX) = ebt MX(at).
Theorem 4
Let X and Y be two random variables with mgf’s Mx(t) and MY(t), respectively. If Mx(t)
= MY(t) for all values of t, then X and Y have the probability distribution. That is, the
mgf uniquely determines the probability distribution of the random variable. The proof
of this theorem is outside the scope of this work, the reader is referred to volume II of this
book.
This theorem will now be used to show that a linear function of a random variable also
has a normal distribution.
Example 31
Suppose that X has distribution N(µ, σ2).
Let Y = aX + b. show that Y also has a Normal distribution with mean aµ + b and
variance a2 σ2. The mgf of Y is
MY(t) = E(ety) = ebt(MX(at))
from theorem I.
1 1 [
MX(t) = E(etx)
-/,W}ðX, -/,W}ðX,
i q( / o i i q(
[
J k J
o i / , ñ,
 2{  2{
k
√ √
ñ
J[
Let / , then MX(t) becomes
J[
kJ À
Á
1 [
i À [ J3/:ó, Áò
I WX / o i WÁò ÀX
i J3/:ò
Âq / o i q
,
Â√2{ J[ √2{ J[
182
õ÷
MX(at) = i À Á /:
, , ,
Substituting at for t, we have
 : ]: :
Hence,
ñ, ,,
W X
¿ / i i À
= ù/i WÀX
/ i ú Á /:
, , ,
2
,
O/ = ' .
IJ À I À
Solution
Á Á Á
Let a = and b = ' , then Y = aX + b.
3 À
Á Á
From Example 312, Y is a normal random variable with mean aµ + b and variance a2σ2.
aµ + b = ' / 0
À À
Substituting for a and b, we see that
Á Á
(i)
aσ = ,/1
2 2 Á,
Á
(ii)
Thus Y is standard normal (N(0, 1)) random variable.
Theorem 5
Let X and Y be independent random variables and Z = X + Y.
Let MX(t), MY(t) and MZ(t) be the mgfs of the random variables X, Y and Z respectively.
Then
MZ(t) = MX(t) MY(t) (4)
Proof:
By definition,
MZ(t) = E(eZt) = E(e(X + Y)t) = E(eXt+eYt) = E(e+Xt) E(eYt) = MX(t) MY(t).
183
If X1, X2,..…, Xn are independent identically distributed with mgf MX(t), then
MZ(t) = [MX(t)]n.
but this is the mgf of a binomially distributed random variable with parameters (n1 + n2,
P).
0- 0,
ûÄ _ È W1 ' _X0- 0, JÄ Å / 0, 1, 2, … )3 = ): n
fÈ WÅX / g
thus, by theorem 2, Z is a binomial random variable and its pdf is given y
0 ii
ii
In general,if X1, X2,…XK are independent binomial random variables with parameters
(n1, p), (n2, p),…(nk, p) respectively, then the moment generating function of Z = X1,
MX(t) = i À- 3/:Á -
µ 2 and variance σ21 + σ22
, ,
MY(t) = i À, 3/:Á ,
, ,
where µ = µ 1 + µ 2, σ2 = σ21 + σ22. Since i À3/: Á is the mgf of N(µ, σ2), hence by
, ,
184
MZ(t) = i Jj- r3J~ s i Jj, r3J~ s / i JW3J~ XWj- j, X / i JW3J~ Xj , h / h3 = h:
Hence,
The mgf of Poisson distribution with parameter λ is i JjW3J~ X , thus Z is a poisson random
variable with mean (parameter) λ = λ1 + λ2. In general, if X1, X2,…,Xn are independent
P(λ1), then
Z = X1 + X2 +… + Xn
is a Poisson random variable with parameter λ = λ1 + λ2 +…+ λn.
Example 36
Let X and Y be independent gamma random variable with parameters (n1, λ) and (n2, λ)
respectively. Show that Z = X + Y is a gamma random variable withparameter (n1 + n2,
λ).
[
h (3 i
0- 0J3 Jjk [
W(X0- J3 i JkWjJX
Solution
I WX / ü Wi I X / o i k q( / o h0- q(
J[ yW)3 X e y W)3 X
h0- [
h0- [
/ o ( 0J3 i JkWjJX q( / o µ0- J3 i J´ qµ
yW)3 X e Wh ' X0-J3 yW)3 X e
But pe µ0- J3 i J´ qµ / y W)X.
[
0-
MX(t) = x .
Hence,
j
jJ
0,
MY(t) = x .
j
jJ
Similarly
0- 0, 0- 0, 0
MZ(t) = x . x / x / x . ) / )3 = ): .
Thus,
j j j j
jJ jJ jJ jJ
This show that Z is a gamma random variable with parameters (n1 + n2, λ).
In general, if X1, X2 ,… + Xk are independent gamma random variables with parameters
(n1, λ), I = 1,2,…k. then Z = X1 + X2 +… + Xk has a gamma distribution with parameters
(n, λ) where n = n1 + n2 +…+ nk.
Example 37
Suppose that Xi has χ2 (n1) (chi-square distribution with ni degrees of freedom) I = 1,
2,…k where the Xi’s are independent random variables.
Let Z = X1 + X2 +… + Xk. Show that Z has distribution χ2n where n = n1 + n2 +…+ nk.
Recall (5.9) that chi-square distribution is a special case of the gamma distribution in
0
which λ = 1/2 and n replaced by n/2. Since the mgf of gamma (n, λ) is given by
x ;
j
jJ
substituting for λ = 1/2 and replacing n by n/2 we have
0/: 0/:
MX(t) = x / x
3/: 3
3/:J 3J:
.
The mgf of Xi is therefore given by
185
0/:
MX(t) / x
3
3J:
I = 1, 2,…
|-|, |
Hence, the mgf of Z is
|
Example 38
σ2. Show that (i) E(^ý ) = µ, where ^ý = (X1 + X2 +…+ Xn) (ii) Var(^ý ) = and
3 Á,
Let X1, X2,… Xn be independent identically normal variables with mean µ and variance
0 0
O/
∑|
- WI- J ÀX
,
Á,
2
(iii) has a χ (n).
O/ / = = =
∑|
(ii)
- WI- JÀX WI- JÀX, WI, JÀX, WI| JÀX,
,
Á, Á, Á, Á,
I- JÀ : I, JÀ : I| JÀ :
(iii)
Example 39
Let X1, X2,…Xn be independent identical normal random variables with mean µ and
variance σ2
Show that : / - ,-
∑| WI JIýX,
Á
has a χ2(n-1).
X1 - µ = Xi - ^ý + ^ý - µ
We shall proceed as follows:
ý : ∑0 ∑0 W ý XW ý X
∑0þ©3WX þ ' μX: / ∑0þ©3 WXþ ' ^ X = þ©3WX ' μX = 2 þ©3 X þ ' ^ ^ ' μ
:
Summing, we have
Á, Á, Á,
186
W J ÀX,
From Example 37, we have ∑0þ©3
Á,
is χ2(n) and (it is left as an exercise to show that
∑W J ÀX,
Á,
is χ2(1).
∑W J ÀX, ∑W J ÀX,
Á, Á,
Let Z = and Y =
∑W J
X,
Á,
2
S =
and S are stochastically independent, we have
Since X 2
|}-
hence,
E(e ) / x
3 ,
3J:
tS 2
and this is the mgf of a chi-square distribution with n – 1 degrees of freedom. This shows
that S2 has a chi-square distribution with n – 1 degrees of freedom.
Exercise 7.3
0 0
2. Let X1, X2,…Xnbe independent identically distributed exponential random
variables with parameters λ. Show that Z = X1 + X2 +…+ Xn has a gamma
distribution with parameters n and λ.
3. Let X1, X2,…Xn be independent N(µ 1, σ21). Show that
Y = a1 X1 + a2X2 +…+ anXn is a normal random variable with mean µ = a1µ 1 +
a2µ 2 +…+ anµ n and variance σ2 = a21σ21 + a2σ22 + …+ a2n σn2.
∑WI JIýX,
0J3
4. Let S2 = where X1, X2,… Xn are independent identically distributed
W0J3Xó ,
Á,
normal random variable with mean µ and variance σ2. Show that has a chi-
square distribution with n – 1 degrees of freedom.
5. If X has χ2(n), show that E(X) = n and Var(X) = 2n.
6. If X –χ2k1 and Y - χ2k1 where X and Y are independent random varaible. Prove that
X + Y - χ2k1+k2.
If X – N(0, 1) prove that (i) Y = X2 has χ2(1)
IýJ À :
If X – N(µ, σ2) prove that H M ' χ: W1X
7. (a)
Á/√0
(b)
∑WI JIýX,
0J3
8. Let S2 = where Xi are independent identically distributed random
:Á 8
0J3
(iii) Show that ^ý and S are stochastically independent.
variables N(µ, σ2), prove that (i) E(S2) = σ2 (ii) Var(S2) = .
2
187
function of (i) Y = -2 In X and (ii) Z = '2 ∑0þ©3 ln ^þ , where the Xi’s are
9. Suppose X is uniformly distributed over (0, 1). Find the moment generating
|
188
There are 3! = 6 permutations of (x1, x2, x3) and hence 6 possible ordered statistics. We
shall now find the joint probability density function of X(1), X(2), X(3). Since for any
» » » » »
N ë(: ' U ^W3X U (: = , (3 = U ^W:X U (3 = , (W>X U (> = ì / » f W(: X
permutation, say (x2< x1< x3)
2 2 2 2 2
» f W(3 X » f W(> X / »> f W(3 XfW(: Xf W(> X,
as illustrated in the Figure (1) below
Since there are 3! Possible permutations, the region D is partitioned into six dejoint sets
D1, D2=,…D6. We have X(1) X(2) X(3).
» » » » »
N ë(3 ' U ^W3X U (3 = U (W:X U (: ' , (> ' U ^W>X U (: ' ì
Thus, for x1< x2< x3
2 2 2 2 2
/ 3! »> f W(3 XWf W(: XXfW(> X
Dividing by ϵ3 and letting ε -> 0, we have
f(x(1), x(2), x(3)) = n! f(x1 f(x2) f(x3); x1< x2< x3
in general, we can extend this idea to a sample of size n and we have for x1 x2 … xn
fx(1), x(2), … x(n) (x1, x2,… xn) = n! f(x1)f(x2)…f(xn);
x1< x2<… < xn.
the marginal density function of X(i) the ith order statistic can be obtained by integrating
1 0 U ( U 1n
f W( X / ¦
Let x1, x2, x3 be a random sample of a random varaible X having the p.d.f.
0 ii
ii
Find
(i) the joint p.d.f. of X1, X2 and X3
(ii) the joint p.d.f of X(1), X(2) and X(3) the ordered statistics
(iii) the marginal p.d.f of X(1).
fk-,k, ,kA W(3 , (: , (> X / fk- W(3 Xfk, W(: XfkA W(> X ' 1; 0 U (3 , (: , (> < 1
Solution
fkW-X,kW,X,kWAX r(W3X , (W:X , (W>X s / 3! fk- W(3 Xfk, W(: XfkA W(> X; 0 U (3 , (: , (> < 1
189
3 3
fk- W(3 X / o o 3! q(W>X
q(W:X
k- k,
3
1 1 3!
/ 3! o W1 ' (: Xq(: / 3! ' (3 ' (3 : / 1 ' 2(3 ' (3 :
k- 2 2 2
/ 3W1 ' (3 X:
0 ii
ii
Another method of finding the joint marginal p.d.f of X(1), X(2),… X(n) is by direct
reasoning as follows:
(i) For X(i) to equal x, I – 1 of (x1, x2….xn) should be less than x and n-1 should be
greater than x and 1 of them equals x.
^' x ^=
: :
P(I – 1 of the Xi’s are less than x) = [F(x)]i-1
P(n – 1 of the Xi’s are greater than x) = [1 – F(x)]n-i
Thus for any given set of partition of (X1, X2…Xn) into 3 as stated above, the required
xi xj
Figure 3
The probability that exactly i of the Xi’s lie in (-∞, x) and (n-1) lie in (x, ∞) is
n
Ci [F(x)]i[1 – F(xi)]n-1
since F(x) = P(X ≤ x) = P(any of the x1, …, Xn falls in (-∞, x).
{(x1, xn) : X(1) ≤ x1, x(n) ≤ x(n)} = {(x1, xn) : X(n) ≤ x(n)}\{x, y)}:
X(1)> x1, x(n) ≤ xn}.
See figure 5 below.
X(n)
X(n)
Figure 5
We know that P(X(n) ≤ xn) = [F(xn)]n, therefore
P(X(1)> x1, X(n)) = P(x1< X1 ≤ xn, x1< x2 ≤ xn… x1< Xn ≤ xn)
= (F(xn) – F(x1))n
Thus
fkW-XkW|X W(3 , (0 X / P(X(1) ≤ x1, X(n) ≤ xn) = [F(xn)]n – {F(xn) - F(x1)}n (12)
Ú:
from which the density function can be derived by differentiation:
fkW-X,kW,X W(3 , (0 X / W( , ( X
Ú(3 Ú(0 kW-XkW|X 3 0
= n(n-1) f(x1)f(xn) [F(xn) – F(x1)]n-2, x1 ≤ xn (13)
Distribution Function of the Range of a Random Sample
The random variable R, defined by R = X(n) – X(1) is called the range of the observed
k- ¡ Wk- ¡XJWk- X
1
Making the change of variable y = F(Xn) – F(x1), dy = f(xn) dxn
191
FR(r) = P(R ≤ r) = npJ[ W(3 = X ' W(3 X0J3 f W(3 Xq(3
Thus
[
FR(r) = n(n - 1) pJ[ W(3 XfW(3 = X W(3 = X' W(3 X0J: q(3
[
(14)
Example 40
Let X1, X2,… Xn be a random sample from uniform distribution on (0, 1). Find
(i) The distribution function of R = X(n) – X(1)
(ii) The probability density function of R.
Where F(x) = pe 1 q( / (;
k
(i)
F(x1 + r) = 1 if x1 + r > 1
= x1 + r if x1 + r < 1.
W XW X 0J: 0 ≤ ≤ 1n
fR(r) = F’R(r) = g) ) ' 1 1 '
FR(r) = n(1 – r)rn-1 + rn
0 Û
ili
(ii) .
Another method of finding the probability density function of R is by using the formula
Solution
fk W( X / , ] U ( U a.
3
The pdf of X is
J
F(X) = p q( /
k 3 kJ
where
J J
( ' ] 0J3 1 )Wa ' (X0J3
fkW-X W(X / ) x1 ' / ,] U ( U a
a'] a'] Wa ' ]X0
,] U ( U a n
0WJkX|}-
kJ 0J3
fkW-X W(X / ) x / \
3
WJX|
ii
ii
J J
(ii)
0
192
Exercise 7.4
1. Let X1, X2,…Xn be independent and each is uniformly distributed over (a, b). find
the joint p.d.f of X(1), X(2),… X(n).
2. Let X1, X2,… Xn be independent random variables each having an exponential
density with parameter λ. Find the pdf of (i) X(1) = min (X1, X2,… Xn)
(ii) Xn = max (X1, X2,… Xn).
3. Let X1, X2,… Xn (n odd) be independent random varaibles each uniformly
distributed over (0, 1). Find the p.d.f of the median of the sample.
Note:When n is odd, the median is ^H|-M.
,
4. Show that if n people are distributed at random along road Y miles long, then the
W0J3X 0
x1 ' , U
probability that no 2 people are less than a distance k miles apart is
¿
w 0J3
5. Let X1, X2,… X2n-1 be a random sample from uniform distribution on (0, 1). Show
that the median of the sample has a beta distribution with parameter (n + 1, n + 1).
193
MODULE FOUR
UNIT 1
LIMIT THEOREMS
PDX-μD m k σ}} L
Or equivalently
3
.
,
Proof:
Let X be a non-negative random variable such that E(X) = µ <∞.
Define another random variable Y as follows:
0 if X < k
Y =
k if X > k
This new variable Y is a discrete variable having two values 0,k. The probability density
function of Y can be written thus:
y 0 k
Hence,
E(Y) =OP(X < k) + kP(X > k)
That is
E(Y) = kP(X ≥ k).
194
Since the variable X ≥ Y for all possible values, we have
E(X) ≥ E(Y) = kP(X ≥ k),
WIX
Thus,
P(X ≥ k) ≤
. (2)
Equation (2) is called the Markov inequality and can be generalized thus: For any j ≥ 0, k
>0
é
P{DWD ≥ k} ≤
(3)
The proof of Chebychev’s inequality is an immediate consequences of (3) by putting j =
/
WIJ KX, Á,
2 and W = X - µ, we obtain
, ,
P{DX - µD ≥ k} ≤
2 2
Since (X - µ) ≥ k <==>DX-µD> k, we have
Á,
,
P{DX-µD> k } = P {(X - µ)2 ≥ k2} ≤ .
Á,
Thus
,
P{DX-µD≤ k } ≤ .
Let X be a random variable having Poisson distribution with mean λ and variance σ2. Use
Example 8.1
j
(i) P{DX-µD ≥ k } ≤ , .
Putting k = 1 we obtain
P{DX-µD ≥ k } ≤ λ.
i 2 / ///Y Â2 /
σ2 4 42
k λ h
On substituting for σ2 we have
(ii)
:
λ2 = 4k2, k2 = H M , / .
: :
j 5
Thus
: j
P{DX-µD ≥ } ≤ .
Since DX-λD ≥ if and only if X-λ <
j Jj
: :
P{DX-λD ≥ } = P{X-λ <- n M = N H^ ' h Y M
or X-λ> λ/2, we see that
j j j
: : :
= P(H^ U M = N H^ Y M ≤
j >j 5
: : j
Since P H^ U M m, i
]Ói
j
:
PH^ Y M ≤ .
>j 5
: j
195
8.2 The Law of Large Numbers of Bernoulli Trails
Let X1, X2,…, Xn be n independent and identically distributed Bernoulli random variables
and let X = X1 + X2 +…+ Xn be Binomial random variables (number of successes), with
parameters n and p. The mean and variable of X are np and npq respectively.
The mean µ grows as n increases but the standard deviation grows only as √), Using
P{DX - npD> ε} ≤ , ≤ ,.
0
Chebychev’s inequality
0
Theorem 8.2
Let X be the number of successes in n independent Bernoulli trails with probability p of
N ¦ ' N Y ¸ ≤ , ≤ ,
success. For any ε > 0.
I 3
0 0 5 0
lim0Z[ N ¦ ' N Y ¸ / 0
I
and
0
Proof:
ü H M / N, ] H M / , . )_ /
3
Applying Chebychev’s inequality
I I
0 0 0 0
±
æ¡H M
|
/ 'Y 0 ] )'Y ∞
», 0»,
_ / W1 ' _ X_ ≤ fÛ ] _, 0 ≤ _ U 1
3
5
Theorem 8.3: A (WLLN) Weak Law of Large Numbers for Independent Random
Variables. Let X1, X2…Xn be independent and identically distributed random variables
ü ¦f H |M¸ Z f W_X.
p ε [0, 1]
ó
0
Corollary II: Weierstress Approximation Theorem.
Let f(X) be a continuous real function on [0, 1] defined on the real interval 0 ≤ p ≤1 and
Theorem 8.5
A WLLN for independent but not necessarily identical random variables.
as n becomes large. The answer to this question is given by the so called central limit
theorem. The theorem asserts that under quite general conditions the sum of independent
variables has the Normal distribution in the limit.
$|
JK
Æ0 / Æ0 /
ó| J 0K |
√0Á, `Á, /0
(i) (ii)
The m.g.f of Zn is
0
È| WX / i W`0K/ÁX xI H M
Á √0
=
WeX ,
:!
M(t) = 1 + M’(0)t + M’’
where A is the remainder term from Maclaurin series. Similarly, when the Maclaurin
series
= =
I, IA
: >
In (1 + x) = x -
we find that
In H1 = = = %M
K WK, Á , X
√0Á 0 Á,
:
= = %- x = = % =
K WK, Á , X 3 K WK, Á, X
√0Á 0 Á, : √0Á 0 Á,
=
¹ = = %¹ U 1.
K WK, Á, X
√0Á :0Á,
199
!
) È| WX / '√)
Â
! W! : ' ! : X 1 ! W! : = Â : X :
:
=) ¨ = = %' = = %
= … &
 √) 2) : 2 Â√) 2) :
) È| WX / .
,
:
Hence,
lim È| WX / i ,
, /:
0Z[
the moment generating function of a random variable having the standard normal
distribution.
Let Xn- n > 1, and X be random variables such that lim0Z[ I' WX / I WX, '∞ U U
Theorem
Example 8.2
A boy throws a fair die 100 times. What is the probability that his mean score will exceed
3.
Let X represent the outcome of each toss.
Possible values of X are 1, 2, 3, 4, 5, 6.
E(X) = 3.5. Var(X) = 2.92.
That is µ = 3.5. σ2 = 2.92
Let X be the mean score. By the central limit theorem X is normally distributed with
mean 3.5 and Varσ2/n.
/ .
Á, :.d:
0 3ee
Hence,
Æ/
IýJ >.4
√e.e:d
200
Example 8.3
Marks in an I.Q examination are normally distributed with mean 55 and standard
deviation 10. What is the probability that
(i) the mean mark of a group of 10 students will be above 50
(ii) the mean mark of a group of 20 students will be between 40 and 50.
(iii) the sum of the marks of the 10 students will be less than 500.
Solution
Let X1, X2,… X10 be the mark scored by each student respectively.
(i) ^ý = mean, be the central limit theory (assuming n = 10 is large enough for this
Iý J 44
Á/√0
distribution) has a standard Normal distribution.
/
IýJ 44 Iý J 44
3e/√3e √3e
(ii) The mean mark of a group of 20 students between will be between 40 and 50. In
/ / / '1.53
ó-. J 44e 4ee J 44e J4e
√3eee 3e√3e 3e√3e
Hence,
P(Sn< 500) = 0.057.
201
8.6 Normal Approximately to the Binomial Distribution
Let X1, X2, …, Xn be n independent identically distributed Bernoulli trails where
1 with probability P
Xi =
0 with probability 1 - P
E(Xi) = p and Var(Xi) = p(1 – p). From theorem 2, the distribution
ó| J Wó| X
√0Á,
--> standard Normal.
where Sn = X1 + X2 + … + Xn
E(Sn) = np
Var(Sn) = np(1 – p), σ2 = p(1 – p).
Note that Sn has binomial distribution with mean np and variance np(1-P).
P(Sn ≤ y) = ∑©e
w n
Ck Pk (1 – p)n-k
For large n
P(Sn ≤ y) = N gÆ0 ≤
w J 0
(
`0W3 J X
)/ H M
w J 0
`0W3 J X
Example 8.3
It is claimed that 60% of the voters I a given ward are going to vote for party A. assuming
that all voters will vote, and that there are 100 voters. What is the probability that Party A
receives at least 50 votes.
202
If a random variable X has a poisson distribution whose mean is λ, then for large λ the
standardized random variable.
IJ j
Æ/
√j
has a standard Normal Distribution. you will recall that E(X) = λ , Var(X) = λ,
Continuity correction may be introduced.
It has been shown that the approximation is good for λ > 5.
Example 8.4
A system suffers random breakdown at a constant rate of 10 per month. Find the
probability that there will be at least 8 breakdowns in any month.
Let X be the number of time the system breaks down in any month. Then X has a poisson
distribution with λ = 10. Thus
NW^ Y 8X / ∑[
~ -. W3eX
©; !
NW^ m 8X / N HÆ Y M / N HÆ Y M / 0.7357
; J 3e J:
√3e √3e
Exercise 8
1. It is claimed that 30% of the voters in a given local government area are going to
vote for party A in a local government election. Assuming that all voters will vote,
there are 4000 voters and the claim is based on a proper (unbiased) sampling
method. What is the probability that Party A will receive more than 1,500 votes.
2. (a) Suppose that a system consists of components each of which has a
probability of 0.05 of failing during a specific time. The system functions
properly when at least 150 components function. Assuming these
components function independently of one another, what is the probability
that the system functions properly during a specific time.
(b) Suppose that the above system is made of n components each having
probability of 0.90 of not failing during a given time. The system will
203
function if at least 80 percent of the components function properly.
Determine n so that the probability that the system functions properly
during a specific time is 0.96.
3. Let X be a non-negative interger-valued random variable whose probability
generating function PX(s) = E(sX) is finite all s.
Use Chebychev’s inequality to verify the following inequalities:
N W^ Y X ≤ . Y 1
*± WòX
ó
(b)
5. Prove that (i) 1 – x ≤ e-x for real x, (ii) log x ≤ x -1 for x > 0, and
(iii) ∏,
þ©3 ^þ ≤ i lf ^þ ≤ 1, Ò / 1,2, …
Ñ
J ∑- I
204