0% found this document useful (0 votes)
3 views44 pages

Stat

Statistics is the branch of mathematics focused on the collection, presentation, analysis, and interpretation of numerical data. It involves steps such as data collection, organization, analysis using mathematical tools, and drawing conclusions. Key concepts include measures of central tendency (mean, median, mode), types of data (qualitative and quantitative), and the importance of statistics in decision-making across various fields.

Uploaded by

impostera618
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
3 views44 pages

Stat

Statistics is the branch of mathematics focused on the collection, presentation, analysis, and interpretation of numerical data. It involves steps such as data collection, organization, analysis using mathematical tools, and drawing conclusions. Key concepts include measures of central tendency (mean, median, mode), types of data (qualitative and quantitative), and the importance of statistics in decision-making across various fields.

Uploaded by

impostera618
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

1. a What do you mean by Statistics?

"Statistics may be defined as the collection,


presentation, analysis, and interpretation of numerical data" – Discuss

Meaning of Statistics

Statistics is the branch of mathematics that deals with numerical data. It helps us to
understand, organize, and draw conclusions from data.

A common definition is:


“Statistics may be defined as the collection, presentation, analysis, and interpretation of
numerical data.”

Now let’s discuss this definition step by step:

1. Collection of Data

This is the first step of statistics. It means gathering information from different sources.

• Example: collecting students’ marks in an exam

• Methods: survey, census, questionnaire, observation, etc.

Without proper data collection, analysis becomes meaningless.

2. Presentation of Data

After collecting data, we need to organize it in a simple and meaningful way.

• Using tables (tabulation)

• Graphs (bar chart, histogram, pie chart)

• Diagrams

This makes the data easy to understand at a glance.

3. Analysis of Data

In this step, we apply mathematical tools to study the data.

• Finding average (mean, median, mode)

• Measuring variation (range, variance, standard deviation)


• Comparing data sets

This helps to extract useful information from raw data.

4. Interpretation of Data

This is the final step where we draw conclusions from the analyzed data.

• Explaining what the results mean

• Making decisions based on data

• Predicting future trends

Example: If average marks are low, we may conclude that students need better preparation.

1. b What are the characteristics of statistics? Discuss.

Characteristics of Statistics

Statistics has some important characteristics that define its nature and scope. These are
discussed below:

1. Aggregates of Facts

Statistics deals with groups of data, not single observation.

• Example: It studies marks of all students in a class, not just one student.
So, individual data is not statistics, but a collection is.

2. Numerically Expressed

Statistics always deals with numerical (quantitative) data.

• Example: height = 160 cm, income = 5000 taka


Qualitative facts like “good”, “bad” must be converted into numbers.

3. Affected by Multiplicity of Causes

Statistical data is influenced by many factors at the same time.


• Example: Student performance depends on study habits, teaching quality, environment,
etc.

So, results are not caused by a single factor.

4. Collected for a Pre-determined Purpose

Data must be collected with a clear objective.

• Example: collecting rainfall data to study climate change


Without purpose, data has no value.

5. Reasonable Accuracy, Not Exactness

Statistics deals with approximate values, not exact results.

• Example: average income or average marks are estimates.

Small errors are acceptable in statistical work.

6. Comparable and Related Data

Statistics must be comparable to be meaningful.

• Example: comparing marks of two classes in the same subject


Without comparability, analysis is not useful.

7. Capable of Being Classified and Tabulated

Statistical data can be organized into tables and categories.

• Example: grouping students by marks range (0–10, 11–20, etc.)

This helps in better analysis.

1. c Describe the importance and uses of statistics.

Importance and Uses of Statistics

Statistics plays a very important role in modern life. It is used in almost every field for decision-
making, analysis, and planning. Its importance and uses are described below:
1. Importance of Statistics

(i) Helps in Decision Making

Statistics provides facts and figures that help people and organizations make correct decisions.

• Example: A company uses sales data to decide production levels.

(ii) Simplifies Complex Data

Large and complex data can be summarized into simple forms using statistics.

• Example: Average marks represent the performance of an entire class.

(iii) Helps in Planning

Governments and organizations use statistics for future planning.

• Example: Population data helps in planning schools, hospitals, and roads.

(iv) Measures Performance

Statistics helps to measure and compare performance.

• Example: Comparing exam results of different schools.

(v) Supports Research Work

Statistics is essential in scientific and social research.

• Example: In medicine, it helps test effectiveness of new drugs.

2. Uses of Statistics

(i) In Economics

• To study income, inflation, unemployment, and GDP

• Helps in economic planning and policy making


(ii) In Business and Industry

• Market research and demand forecasting

• Profit analysis and cost control

• Quality control in production

(iii) In Government and Administration

• Population census

• Planning national budgets

• Policy formulation (education, health, etc.)

(iv) In Education

• Analysis of student performance

• Examination results and grading systems

• Educational research

(v) In Medicine and Health Science

• Study of diseases and treatment effectiveness

• Hospital management

• Public health planning (vaccination programs, etc.)

(vi) In Banking and Insurance

• Risk estimation

• Loan analysis

• Premium calculation in insurance

1. d Define with examples: (i) Population (ii) Sample (iii) Variable (iv) Constant
Here are the definitions with simple examples:

1. Population
Population means the entire group of people, objects, or observations about which we
want information.

Example:
All students of a university are a population.

2. Sample
A sample is a part or subset taken from the population for study.

Example:
100 students selected from the university for a survey are a sample.

3. Variable
A variable is a characteristic that can change or take different values.

Example:
Height, age, marks, and income are variables because their values differ from person to person.

4. Constant
A constant is a quantity or characteristic that remains fixed and does not change.

Example:
In a class, if every student studies in the same university, then the university name is a constant.

1. e What is a variable? Define its types with examples.\

What is a Variable?

A variable is a characteristic or attribute that varies from one unit to another within a
population. For example, if you are studying a group of students, "height" is a variable because
it differs from person to person.

Classification of Variables
Variables are broadly classified into two categories based on how they are measured or
described:

1. Qualitative Variable

These describe attributes or qualities that cannot be expressed numerically. They represent
categories or labels.

• Examples: Hair color, gender, religion, or a father's occupation.

2. Quantitative Variable

These are characteristics that can be expressed in numerical form or counted.

• Example: Monthly family income or age.

Quantitative variables are further divided into two sub-groups:

• Discrete Variable: A variable that takes on isolated or whole numbers (integers). There
are no values in between the points.

o Examples: Number of siblings, family size, or the number of road accidents per
day.

• Continuous Variable: A variable that can take any value within a specific range or limit,
including decimals and fractions.

o Examples: Height (e.g., 170.5 cm), weight, or temperature.

1. f What are origin and scale? Discuss their uses.

Origin

The origin is the starting or reference point from which measurements are taken.
It indicates the zero point of a scale.

Example:
In Celsius temperature scale, 0°C is an origin.

Scale

The scale is the unit or proportion used to measure observations.


It shows how much distance on the measuring instrument represents a unit value.

Example:
In a ruler, each centimeter division represents a fixed scale.
Uses of Origin and Scale

1. Measurement of Data
They help in measuring and expressing numerical data properly.

2. Comparison
They make comparison between observations easy and meaningful.

3. Classification and Analysis


Statistical data can be classified and analyzed accurately using proper origin and scale.

4. Graphical Representation
Graphs and charts are drawn correctly by choosing suitable origin and scale.

5. Transformation of Data
They are used in changing units or simplifying calculations in statistics and mathematics.

1. g Discuss the different types of data.

1. Qualitative Data

This refers to information that cannot be measured in numerical form. It describes qualities,
attributes, or categories rather than quantities.

• Examples from the text: Religion, economic condition, color, and gender.

• Key characteristic: It answers questions like "what kind" or "which type."

2. Quantitative Data

This refers to information that can be expressed in numerical form or numbers. It represents
counts or measurements.

• Examples from the text: Family size, population size, height, weight, and monthly
income.

• Key characteristic: It answers questions like "how many," "how much," or "how often."

Quick Comparison

Feature Qualitative Data Quantitative Data

Nature Descriptive / Non-numerical Numerical / Measurable


Feature Qualitative Data Quantitative Data

Focus Qualities and attributes Quantities and amounts

Examples Eye color, City of birth Age, Temperature, Test scores

1. h Define measurement and measurement scales. Explain types with examples.

Measurement is the process of assigning numbers or symbols to the characteristics of objects or


individuals according to certain rules.

Measurement scales are the systems or methods used to classify and measure data.

1. i Define graphical representation. Describe different frequency distribution graphs.

1. i Define graphical representation. Describe different frequency distribution graphs.

Graphical Representation

Graphical representation means showing data by graphs or charts so that it becomes easy to
understand.
Example:
Showing students’ marks by a graph.

Types of Frequency Distribution Graphs

1. Histogram

A histogram is made with rectangles joined together.

• X-axis → class intervals

• Y-axis → frequencies

Example:
Marks: 0–10, 10–20, 20–30, etc.

Use:
Shows how data are distributed.

2. Frequency Polygon

It is made by joining the middle points of histogram bars with straight lines.

Use:
Easy to compare data.

3. Frequency Curve

A smooth curve drawn instead of straight lines in a frequency polygon.

Use:
Shows the general shape of data.

4. Ogive (Cumulative Frequency Curve)

An ogive shows cumulative frequencies.

Two types:

• Less than ogive


• More than ogive

Use:
Used to find median and quartiles.

2. a Discuss the concepts of central tendency in statistics.

Central Tendency in Statistics

Central tendency means the value that represents the center or typical value of a group of data.
It shows around which value most of the data are gathered.

In simple words, it is a method of finding the “average” value of data.

Examples of central tendency:

• Mean

• Median

• Mode

1. Arithmetic Mean (Average)

The arithmetic mean is obtained by dividing the sum of all observations by the total number of
observations.

Formula
∑𝑥
𝑥ˉ =
𝑛
Where,

• x = observations

• n = number of observations

Example

For data: 2, 4, 6
Mean = (2 + 4 + 6)/3 = 4

Uses

• Most commonly used average

• Easy to calculate
2. Median

Median is the middle value of a dataset arranged in ascending or descending order.

Example

For data: 2, 4, 6, 8, 10
Median = 6

Uses

• Useful when data contain extreme values

• Suitable for skewed data

3. Mode

Mode is the value that occurs most frequently in a dataset.

Example

For data: 2, 3, 3, 5, 7
Mode = 3

Uses

• Useful for qualitative data

• Shows the most common value

2.b What do you mean by measures of central tendency? Describe the various
measures of central tendency with their merits, demerits, and uses.

What are Measures of Central Tendency?

A Measure of Central Tendency is a single value that attempts to describe a set of data by
identifying the central position within that set of data. They are also known as measures of
location or center.

1. Arithmetic Mean (AM)

The Arithmetic Mean is the sum of all observations divided by the total number of observations.

Formula: $\bar{x} = \frac{\sum x}{n}$


• Merits: Easy to calculate and understand; utilizes every value in the dataset.

• Demerits: Highly affected by outliers (extreme values); not suitable for qualitative data.

• Uses: Used in schools for average marks, economics for per capita income, and general
statistics.

2. Geometric Mean (GM)

The Geometric Mean is the $n^{th}$ root of the product of $n$ observations.

• Merits: Less affected by extreme values than AM; useful for calculating ratios and
percentages.

• Demerits: Cannot be calculated if any value is zero or negative; mathematically complex.

• Uses: Used to find the average growth rate (population growth) or compound interest.

3. Harmonic Mean (HM)

The Harmonic Mean is the reciprocal of the arithmetic mean of the reciprocals of the data
values.

• Merits: Gives more weight to smaller values; well-defined by a mathematical formula.

• Demerits: Difficult to calculate; cannot be used if any value is zero.

• Uses: Ideal for finding the average speed or rates involving time and distance.

4. Median

The Median is the middle value of a dataset when the observations are arranged in ascending or
descending order.

• Merits: Not affected by extreme values; easy to find in a ranked list.

• Demerits: Does not use all observations (it only cares about the middle position); not
suitable for further algebraic treatment.

• Uses: Used for skewed data, such as calculating the median household income.

5. Mode

The Mode is the value that appears most frequently in a dataset.

• Merits: Very easy to identify; the only measure that can be used for qualitative data
(e.g., favorite color).
• Demerits: A dataset can have no mode or multiple modes (bimodal/multimodal); it is
not based on all observations.

• Uses: Used in business and market research to identify the most popular product or
size.

2. c What are the essential characteristics of an ideal average? Which measure of central
tendency is the best and why?

Essential Characteristics of an Ideal Average

An ideal average should have the following characteristics:

1. Easy to Understand
It should be simple and clear.

2. Easy to Calculate
Its calculation should not be difficult.

3. Based on All Observations


It should consider all values of the data.

4. Not Affected by Extreme Values


Very large or very small values should not affect it much.

5. Rigidly Defined
Its value should be definite and unique.

6. Suitable for Further Mathematical Treatment


It should help in further statistical calculations.

7. Stable in Sampling
Small changes in data should not change it greatly.

Best Measure of Central Tendency

The Arithmetic Mean is considered the best measure of central tendency.

Reasons

• It uses all observations.

• It is easy to calculate and understand.


• It is rigidly defined.

• It is suitable for algebraic and statistical analysis.

• It is widely used in mathematics, economics, and statistics.

However, when data contain extreme values, the median may be more suitable

2. d Critically compare the various measures of central tendency.

2. e Describe the calculation method of median and mode from the graph.

Based on the images you provided from your textbook, here is a detailed description of how to
determine the Median and Mode graphically.

1. Determination of Median from Graph (Ogive Curve)

The median is found by plotting a Cumulative Frequency Curve, commonly known as an Ogive.

Step-by-Step Method:

1. Preparation: First, calculate the cumulative frequencies from the given frequency
distribution.
2. Plotting: * Plot the Upper Class Limits along the X-axis.

• Plot the corresponding Cumulative Frequencies along the Y-axis.

3. Drawing the Curve: Join all the plotted points with a smooth freehand curve to obtain
the Ogive curve (labeled AB in your image).

4. Finding the Median Position: Identify the point on the Y-axis that represents N/2 (where
N is the total number of observations).

5. Intersection: From the N/2 position on the Y-axis, draw a horizontal line parallel to the X-
axis until it intersects the Ogive curve at Point C.

6. Final Result: Draw a perpendicular line (CM) from Point C down to the X-axis. The
distance OM on the X-axis indicates the Median value

2. Graphical Determination of Mode (Histogram)

The mode is determined using a Histogram, specifically by looking at the highest rectangle
which represents the Modal Class.

Step-by-Step Method:

1. Draw the Histogram: Represent the frequency distribution as a histogram with class
intervals on the X-axis and frequency on the Y-axis.

2. Identify the Modal Class: Locate the highest rectangle in the histogram.

3. Drawing Connecting Lines: To pinpoint the mode within the highest bar:
o Draw a straight line from the top-left corner of the highest bar (A) to the top-left
corner of the next adjacent bar (C).

o Draw another straight line from the top-right corner of the highest bar (B) to the
top-right corner of the previous adjacent bar (D).

4. Find the Intersection: These two lines will intersect at a specific point (Q).

5. Final Result: Draw a perpendicular line from the intersection point Q down to the X-axis
(Point P). The value at point P (the abscissa) on the X-axis is the Mode.

2. f When is the median more preferable to the arithmetic mean?

Median is more preferable than arithmetic mean in the following situations:

1. When Extreme Values Exist


If the data contain very large or very small values, median is better because it is not
affected by them.

Example:
Income distribution where a few people are extremely rich.

2. For Skewed Data


When data are not symmetrical, median gives a better central value.
3. For Open-End Class Intervals
Median can be calculated even when class intervals are open-ended.

Example:
“50 and above”.

4. For Qualitative Ranking Data


Median is suitable for ordered data like grades or ranks.

5. When Data are Incomplete or Irregular


Median is useful when some observations are missing or uncertain.

2. g Establish the relationship between arithmetic mean, geometric mean, and harmonic mean.

2. h When are the arithmetic mean, geometric mean, and harmonic mean equal?
3. a Discuss the concept of correlation. Write down the importance and uses of correlation.

Concept of Correlation

Correlation is a statistical tool used to measure the strength and direction of the relationship
between two variables. It determines how closely two variables move together.

Types:

• Positive: Both variables increase or decrease together (e.g., Height and Weight).

• Negative: One increases while the other decreases (e.g., Price and Demand).

• Zero: No relationship exists between the variables.

Importance & Uses

1. Measuring Association: It quantifies how strongly two variables are linked, helping to
simplify complex data.

2. Basis for Prediction: It is the foundation for Regression Analysis, allowing us to estimate
unknown values based on known data.

3. Business Decisions: Helps managers understand the impact of specific factors, like how
advertising affects sales volume.

4. Policy Making: Governments use it to study social links, such as the correlation between
education levels and poverty.

5. Scientific Research: Used to validate experiments by checking if changes in a variable


lead to expected results.

6. Reliability Check: Helps in verifying the consistency and accuracy of collected data.

3. b Describe the different types of simple correlation.

Different Types of Simple Correlation

1. Perfect Positive Correlation

• Definition: When both variables move in the same direction at an equal rate.

• Value: r = +1

• Example: Radius and circumference of a circle.

2. Partial Positive Correlation


• Definition: When variables move in the same direction, but the rate of change is not
equal.

• Value: 0 < r < 1

• Example: Income and expenditure.

3. Perfect Negative Correlation

• Definition: When variables move in opposite directions at an equal rate.

• Value: r = -1

• Example: Pressure and volume of a gas.

4. Partial Negative Correlation

• Definition: When variables move in opposite directions, but the rate of change is not
equal.

• Value: -1 < r < 0

• Example: Price and demand of a product.

5. Zero Correlation

• Definition: When there is no relationship between the two variables.

• Value: r = 0

• Example: Height and intelligence.

3. c What is scatter diagram? Discuss the different nature of correlation with the help of scatter
diagram.

A Scatter Diagram (or Scatter Plot) is a graphical method used to study the relationship between
two variables. In this method, the values of two variables are plotted as points (dots) on a graph
sheet with an X-axis and a Y-axis. By observing the pattern or "scatter" of these points, we can
visually identify the nature and strength of the correlation between the variables.

3. d Explain the properties of Pearson’s coefficient of correlation.

Properties of Pearson’s Coefficient of Correlation

Pearson’s coefficient of correlation measures the degree of linear relationship between two
variables. It is denoted by r.

−1 ≤ 𝑟 ≤ +1
Properties

1. Value Lies Between −1 and +1

• r = +1 → perfect positive correlation

• r = −1 → perfect negative correlation

• r = 0 → no correlation

2. Indicates Direction of Relationship

• Positive value → variables move in the same direction

• Negative value → variables move in opposite directions

3. Measures Linear Relationship

Pearson’s correlation measures only linear correlation between variables.

4. Unit Free Measure

It has no unit because it is independent of the units of measurement.

5. Symmetric Property

Correlation between X and Y is the same as between Y and X.

𝑟𝑥𝑦 = 𝑟𝑦𝑥

6. Independent of Change of Origin and Scale

Pearson’s coefficient is not affected by change of origin and scale.

7. Uses All Observations

It is calculated using all data values, so it is more reliable.


8. Affected by Extreme Values

Very large or very small observations can affect the value of r.

3. e What is measures of dispersion? Why is it necessary to measure dispersion?

Measures of Dispersion

Measures of dispersion are the statistical methods used to show how much the data values are
spread out or scattered around a central value (average).

In simple words, dispersion tells us whether the data are close together or widely spread.

Examples of measures of dispersion:

• Range

• Quartile Deviation

• Mean Deviation

• Standard Deviation

Necessity of Measuring Dispersion

1. To Test the Reliability of an Average

An average alone may be misleading if the data are widely scattered.


Dispersion shows whether the average properly represents the data.

2. To Compare Variability

Two datasets may have the same average but different spreads.
Measures of dispersion help compare their consistency.

3. To Control Variations

In manufacturing, business, and medicine, dispersion helps detect unwanted variations and
maintain quality.
4. Basis for Further Statistical Analysis

Many statistical methods like correlation, regression, and hypothesis testing depend on
measures of dispersion.

5. To Understand the Nature of Data

Dispersion helps us know whether the data are homogeneous (similar) or heterogeneous
(different).

3. f Discuss the importance and necessity of measuring dispersion.

3. g Describe the various measures of dispersion. Discuss the following measures of dispersion
indicating their relative merits and demerits:

(i) Range

(ii) Mean deviation

(iii) Standard deviation

(iv) Quartile deviation

Measures of Dispersion

Measures of dispersion show how much the data values are spread around the average.

Main measures of dispersion are:

1. Range

2. Mean Deviation

3. Standard Deviation

4. Quartile Deviation

(i) Range

Range is the difference between the largest and smallest values.

𝑅𝑎𝑛𝑔𝑒 = 𝐿 − 𝑆

Where,

• L = largest value
• S = smallest value

Merits

• Very simple to calculate

• Easy to understand

Demerits

• Uses only two observations

• Highly affected by extreme values

• Not reliable

(ii) Mean Deviation

Mean deviation is the average of the absolute deviations of observations from a central value.

Merits

• Uses all observations

• Easy to understand

Demerits

• Ignores algebraic signs

• Not suitable for advanced mathematical analysis

(iii) Standard Deviation

Standard deviation is the square root of the average of squared deviations from the arithmetic
mean.

∑(𝑥 − 𝑥ˉ )2
𝜎=√
𝑁

Merits

• Uses all observations

• Most reliable and widely used


• Suitable for further statistical analysis

Demerits

• Calculation is comparatively difficult

• Affected by extreme values

(iv) Quartile Deviation

Quartile deviation is half of the difference between the third quartile and first quartile.
𝑄3 − 𝑄1
𝑄. 𝐷. =
2
Merits

• Not much affected by extreme values

• Simple to calculate

Demerits

• Uses only part of the data

• Not suitable for detailed analysis

3. h State the properties of standard deviation.

Properties of Standard Deviation (S.D.)

1. Independent of origin
Adding or subtracting the same number from all observations does not change S.D.

2. Dependent on scale
Multiplying or dividing all observations changes S.D. in the same ratio.

3. For two unequal values


S.D. is half of the range.
𝑅𝑎𝑛𝑔𝑒
𝑆. 𝐷. =
2
4. Mean deviation is never greater than S.D.
Standard deviation is always equal to or greater than mean deviation.

𝑀. 𝐷. ≤ 𝑆. 𝐷.

5. For first n natural numbers

𝑛2 − 1
𝑆. 𝐷. = √
12

6. For two positive values


Arithmetic mean is always greater than S.D.

7. S.D. is smaller than range


Standard deviation can never be greater than the range.

𝜎<𝑅

8. S.D. is always positive


It measures the spread of data and cannot be negative.

5. a Define the following terms with examples: experiment, random experiment, event,
composite event, probability.

1. Experiment
An experiment is an action that gives a result.

Example: Tossing a coin or rolling a die.

2. Random Experiment
A random experiment is an experiment where the result cannot be known before it
happens.

Example: Rolling a die. The result may be 1, 2, 3, 4, 5, or 6.

3. Event
An event is a result or group of results of a random experiment.

Example: Getting an even number when rolling a die → {2, 4, 6}.


4. Composite (Compound) Event
A composite event is an event with more than one outcome.

Example: Getting a number greater than 4 when rolling a die → {5, 6}.

5. Probability
Probability means the chance that an event will happen.
Its value is between 0 and 1.

Formula:
Favorable outcomes
𝑃(𝐸) =
Total outcomes
Example:
Probability of getting an Ace from 52 cards = 4/52 = 1/13.

5. b Write down the difference between a priori and a posteriori definition of probability.

5. c Define the conditional probability with example.


5. d What do you mean by permutation and combination with examples?

Permutation and Combination are the two fundamental ways we count and arrange objects in
mathematics. The simplest way to distinguish them is to remember that in Permutations,
order matters, while in Combinations, order does not matter.

1. Permutation (Arrangement)

A permutation is an arrangement of objects in a specific order. If you change the order of the
objects, you create a new permutation.

Key Idea: Order is everything.

Formula:

𝒏
𝒏!
𝑷𝒓 =
(𝒏 − 𝒓)!

Real-life Example: A Phone Passcode


If your passcode is 1-2-3-4, entering 4-3-2-1 will not unlock the phone. Even though the
numbers are the same, the order is different, making it a different permutation.

Math Example:
How many ways can you arrange the letters in the word "CAT"?

Possible ways:
CAT, CTA, ACT, ATC, TCA, TAC

There are 6 different permutations.

2. Combination (Selection)

A combination is a selection of objects where the order does not matter. We are only
interested in which items are picked, not the sequence in which they are picked.
Key Idea: Order is irrelevant.

Formula:

𝒏
𝒏!
𝑪𝒓 =
𝒓! (𝒏 − 𝒓)!

Real-life Example: A Fruit Salad


If you make a salad with Apple, Banana, and Grape, it is the same salad if you put the Grape,
Banana, and then the Apple. The group of fruits is what matters, not the order they went into
the bowl.

Math Example:
If you have 3 friends (A, B, and C) and you can only invite 2 of them to dinner, how many
groups can you form?

Possible groups:
{A, B}, {A, C}, {B, C}

Note that {A, B} is the same as {B, A}, so we only count it once. There are 3 combinations.

5. e Prove the additive law of probability with statement for two mutually disjoint events.
5. f State and prove the additive law of probability for two non-mutually exclusive events.

5. g Explain the difference approaches of probability.

Different Approaches of Probability

There are mainly three approaches of probability:

1. Classical Approach

In this approach, probability is found by dividing favorable outcomes by total possible


outcomes, when all outcomes are equally likely.

Formula:
Favorable outcomes
𝑃(𝐸) =
Total possible outcomes
Example:
When a die is rolled, probability of getting 2 = 1/6.
2. Empirical (Statistical) Approach

In this approach, probability is determined from past data or repeated experiments.

Formula:
Number of times the event occurs
𝑃(𝐸) =
Total number of trials
Example:
If it rains on 20 days out of 100 days, then probability of rain = 20/100 = 0.2.

3. Axiomatic Approach

In this approach, probability is defined by some basic rules or axioms.

Main axioms:

1. Probability is always between 0 and 1.

2. Probability of the sample space is 1.

3. For mutually exclusive events,


the probability of their union is the sum of their probabilities.

Formula:

𝑃(𝐴 ∪ 𝐵) = 𝑃(𝐴) + 𝑃(𝐵)

Example:
If P(A) = 0.3 and P(B) = 0.4, then
P(A ∪ B) = 0.7.

6. a Define the following terms with examples: mutually exclusive events, exhaustive events,
complementary events, sample space, favourable outcomes

1. Mutually Exclusive Events

Definition:
If one event happens, the other event cannot happen at the same time. These are called
mutually exclusive events.
Example:
When a coin is tossed, “Head” and “Tail” cannot come together.
So, Head and Tail are mutually exclusive events.

2. Exhaustive Events

Definition:
All possible outcomes of an experiment together are called exhaustive events.

Example:
When a die is thrown, the possible outcomes are:
{1, 2, 3, 4, 5, 6}
These six outcomes are exhaustive events.

3. Complementary Events

Definition:
If event A does not happen, it is called the complement of A.
It is written as Ā or Aᶜ.

The sum of their probabilities is always:

𝑃(𝐴) + 𝑃(𝐴‾) = 1

Example:
If A = “Bangladesh wins the match”,
then Ā = “Bangladesh does not win the match”.

4. Sample Space

Definition:
The set of all possible outcomes of a random experiment is called the sample space.
It is usually written by S.

Example:
For throwing a die:

S = {1, 2, 3, 4, 5, 6}

Here, S is the sample space.


5. Favourable Outcomes

Definition:
The outcomes which help an event happen are called favourable outcomes.

Example:
A box has 10 balls and 5 are red.
If you want a red ball, then the favourable outcomes are 5.

6.b Prove the multiplicative law of probability with statement for two independent events.

6. c Prove the multiplicative law of probability with statement for two dependent events.
6.d Prove the Bayes’ formula with statement.

Bayes’ Theorem খুব সহজে বুঝজে হজে আজে Conditional Probability বুঝজে হজব।

দুইটা formula নিই:


𝑃(𝐴 ∩ 𝐵)
𝑃(𝐴 ∣ 𝐵) =
𝑃(𝐵)

এবং
𝑃(𝐴 ∩ 𝐵)
𝑃(𝐵 ∣ 𝐴) =
𝑃(𝐴)
এখাজি দুজটাজেই 𝑃(𝐴 ∩ 𝐵)আজে।

এখি নিেীয় formula থেজে পাই:

𝑃(𝐴 ∩ 𝐵) = 𝑃(𝐵 ∣ 𝐴) × 𝑃(𝐴)

এই মািটা প্রেম formula-থে বসাই:

𝑃(𝐵 ∣ 𝐴) × 𝑃(𝐴)
𝑃(𝐴 ∣ 𝐵) =
𝑃(𝐵)

এটাই হজো Bayes Formula.

6. e If A and B are two independent events then prove that 𝐴 and 𝐵 ̅ independent.

6. f The sum of the probabilities of happening and non-happening of an event is one.


7. a Write the difference between probability function and probability density function.

Difference between Probability Function and Probability Density Function

Probability Function (P.F.) Probability Density Function (P.D.F.)

It is used for discrete random variables. It is used for continuous random variables.

It gives the probability of an exact value. It gives the density of probability over an interval.

Written as P(X = x). Written as f(x).

Probability of a value can be directly found. Probability at a single point is zero.

Sum of all probabilities is 1. Total area under the curve is 1.

Example: Tossing a coin, throwing a die. Example: Height, weight, temperature.

7. b What is random variable? How many types of random variables?

Random Variable

Definition

A random variable is a variable that takes different numerical values according to the outcomes
of a random experiment.
It is usually represented by X, Y, Z etc.

Example

When a die is thrown:

• If 1 comes, X = 1

• If 2 comes, X = 2

• …

• If 6 comes, X = 6

Here, X is a random variable.

Types of Random Variables

There are mainly two types of random variables:

1. Discrete Random Variable

A random variable that takes countable values is called a discrete random variable.
Example

• Number of students in a class

• Outcomes of a die: {1,2,3,4,5,6}

2. Continuous Random Variable

A random variable that can take any value within an interval is called a continuous random
variable.

Example

• Height of students

• Weight of a person

• Temperature

So, random variables are of two types:

1. Discrete Random Variable

2. Continuous Random Variable

7. c How can you calculate the expected value of a random variable?

Expected Value of a Random Variable

Definition

The expected value (or mean) of a random variable is the average value we expect from a
random experiment in the long run.

It is denoted by E(X).

Formula for Discrete Random Variable

𝐸(𝑋) = ∑𝑥𝑃(𝑥)

Where:

• x = value of the random variable

• P(x) = probability of x

Steps to Calculate Expected Value

1. Multiply each value of the random variable by its probability.


2. Add all the results.

Example

Suppose:

X P(X)

1 0.2

2 0.5

3 0.3

Then,

𝐸(𝑋) = (1 × 0.2) + (2 × 0.5) + (3 × 0.3)


= 0.2 + 1.0 + 0.9
= 2.1

So, the expected value is 2.1.

7. d Write down the rules and laws of expectation.

1. Expectation of a Constant

The expectation of a constant is the constant itself.

𝐸(𝑐) = 𝑐

Where c is a constant.

2. Constant Multiplication Rule

If a random variable is multiplied by a constant, the expectation is also multiplied by that


constant.

𝐸(𝑐𝑋) = 𝑐𝐸(𝑋)

3. Addition Rule

The expectation of the sum of two random variables equals the sum of their expectations.

𝐸(𝑋 + 𝑌) = 𝐸(𝑋) + 𝐸(𝑌)


4. Subtraction Rule

The expectation of the difference of two random variables equals the difference of their
expectations.

𝐸(𝑋 − 𝑌) = 𝐸(𝑋) − 𝐸(𝑌)

5. Expectation of Product of Independent Variables

If X and Y are independent random variables, then

𝐸(𝑋𝑌) = 𝐸(𝑋)𝐸(𝑌)

6. Expectation of a Linear Function

For constants a and b,

𝐸(𝑎𝑋 + 𝑏) = 𝑎𝐸(𝑋) + 𝑏

These are the basic rules and laws of expectation used in probability and statistics.

7. e Define the variance of the random variable. State the properties of variance of random
variable

Definition

Variance is a measure of dispersion or spread of a random variable from its mean (expected
value).

It shows how far the values are spread from the average value.

The variance of a random variable X is denoted by Var(X) or σ².

Formula of Variance

𝑉𝑎𝑟(𝑋) = 𝐸[(𝑋 − 𝐸(𝑋))2 ]

Another important formula:

𝑉𝑎𝑟(𝑋) = 𝐸(𝑋 2 ) − [𝐸(𝑋)]2

𝑉𝑎𝑟(𝑋) = 𝜎 2 ≈ 1.96
μ-σ+σVar(X) ≈ 1.96

Properties of Variance

1. Variance of a Constant is Zero

𝑉𝑎𝑟(𝑐) = 0

where c is a constant.

2. Variance is Always Non-negative

𝑉𝑎𝑟(𝑋) ≥ 0

Variance can never be negative.

3. Variance of a Constant Multiple

𝑉𝑎𝑟(𝑐𝑋) = 𝑐 2 𝑉𝑎𝑟(𝑋)

where c is a constant.

4. Adding a Constant Does Not Change Variance

𝑉𝑎𝑟(𝑋 + 𝑐) = 𝑉𝑎𝑟(𝑋)

5. Variance of Sum of Independent Variables

If X and Y are independent random variables,

𝑉𝑎𝑟(𝑋 + 𝑌) = 𝑉𝑎𝑟(𝑋) + 𝑉𝑎𝑟(𝑌)

6. Variance of Difference of Independent Variables

If X and Y are independent,

𝑉𝑎𝑟(𝑋 − 𝑌) = 𝑉𝑎𝑟(𝑋) + 𝑉𝑎𝑟(𝑌)


7. f Differentiate between the arithmetic mean and mathematical expectation.

Difference between Arithmetic Mean and Mathematical Expectation

Arithmetic Mean Mathematical Expectation

It is the average value of a random variable based


It is the average of observed data.
on probability.

Used in statistics for actual data. Used in probability theory.

Calculated from possible values and their


Calculated from given observations.
probabilities.

Formula: Total of observations ÷ Number of


Formula: E(X) = ΣxP(x)
observations

Denoted by x̄ (x-bar). Denoted by E(X).

Depends on collected data. Depends on probability distribution.

7. g What is a joint probability function? What conditions must a function satisfy to quality as a
joint probability function?

Definition

A joint probability function gives the probability of two discrete random variables occurring
together.

If X and Y are two random variables, then the joint probability function is written as:

𝑃(𝑋 = 𝑥, 𝑌 = 𝑦) = 𝑓(𝑥, 𝑦)

It means the probability that X = x and Y = y occur at the same time.

Example

Suppose:

• X = number on first die

• Y = number on second die

Then,
1
𝑃(𝑋 = 1, 𝑌 = 2) =
36

because there are 36 possible outcomes when two dice are thrown.

Conditions of a Joint Probability Function

A function must satisfy the following conditions to be a valid joint probability function:

1. Non-negative Condition

The probability can never be negative.

𝑓(𝑥, 𝑦) ≥ 0

for all x and y.

2. Total Probability Condition

The sum of all joint probabilities must be 1.

∑ ∑ 𝑓(𝑥, 𝑦) = 1
𝑦
𝑥

These two conditions are necessary for a function to qualify as a joint probability function.

7. h What do you mean by moment generating function and state its uses, limitations and
properties?

Definition

The moment generating function (MGF) of a random variable is a function used to find the
moments such as mean, variance, etc. of a probability distribution.

It is denoted by Mₓ(t).

Formula of MGF

𝑀𝑋 (𝑡) = 𝐸(𝑒 𝑡𝑋 )

For a discrete random variable,

𝑀𝑋 (𝑡) = ∑𝑒 𝑡𝑥 𝑃(𝑥)
Uses of MGF

1. Finding Mean

MGF helps to find the mean of a random variable.

2. Finding Variance

It is used to calculate variance and higher moments.

3. Identifying Distribution

Different probability distributions have different MGFs, so it helps identify distributions.

4. Simplifying Calculations

It makes many probability calculations easier.

Limitations of MGF

1. May Not Exist

MGF does not exist for all probability distributions.

2. Difficult Calculation

Sometimes MGF is difficult to calculate.

3. Complex for Large Problems

For complicated distributions, the MGF may become complicated.

Properties of MGF

1. Value at Zero

𝑀𝑋 (0) = 1

2. Mean from MGF

The first derivative of MGF at t = 0 gives the mean.

𝑀𝑋′ (0) = 𝐸(𝑋)


3. Second Moment from MGF

The second derivative gives the second moment.

𝑀𝑋′′ (0) = 𝐸(𝑋 2 )

4. Unique Property

Different random variables have different MGFs.

5. MGF of Sum of Independent Variables

If X and Y are independent,

𝑀𝑋+𝑌 (𝑡) = 𝑀𝑋 (𝑡)𝑀𝑌 (𝑡)

You might also like