0% found this document useful (0 votes)
243 views7 pages

Key Concepts in Business Statistics

1. This document provides examples of calculating various statistical measures including the mean, median, mode, standard deviation, and coefficient of variation from individual data points, discrete data series, and continuous data series. 2. It also includes examples of calculating the coefficient of correlation from bivariate data to measure the relationship between two variables. 3. Standard deviation, mean, and coefficient of variation are calculated for various distributions to measure the spread and variability of data.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
243 views7 pages

Key Concepts in Business Statistics

1. This document provides examples of calculating various statistical measures including the mean, median, mode, standard deviation, and coefficient of variation from individual data points, discrete data series, and continuous data series. 2. It also includes examples of calculating the coefficient of correlation from bivariate data to measure the relationship between two variables. 3. Standard deviation, mean, and coefficient of variation are calculated for various distributions to measure the spread and variability of data.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd

BUSINESS STATISTICS

UNIT 1

1.1 Problems on Mean, Median, Mode, SD and CV

Mean- Individual Observations


[Link] following table gives the monthly income of 10 employees in an office:

Income ( Rs) 14,780 15,760 26,690 27,750 24,840


24,920 16,100 17,810 27,050 26,950

Calculate the arithmetic mean of income.

Mean- Discrete Series


2. From the following data of the marks obtained by 60 students of a class, calculate the
arithmetic mean.
Marks 20 30 40 50 60 70
Number of students 8 12 20 10 6 4

Mean- Continuous Series


3. From the following data compute Arithmetic mean by direct method
Marks 0-10 10-20 20-30 30-40 40-50 50-60
[Link] students 5 10 25 30 20 10

4. From the following data compute Arithmetic mean


Marks 0-10 10-30 30-60 60-100 100-150
[Link] students 5 12 20 8 5

Correcting Incorrect Values


[Link] mean marks of 100 students were found to be 40. Later on, it was discovered that a score
of 53 was misread as 83. Find the correct mean corresponding to the correct score.

Combined Mean
6. The mean height of 25 male workers in a factory is 61 inches and the mean height of 35
female workers in the same factory is 58 inches. Find the combined mean height of 60 workers
in the factory.

Median- Individual Observations


7. From the following data of the wages of 7 workers, compute the median wage:

Wages (in Rs) : 14,100 14,150 16,080 17,120 15,200 16,160 17,400

8. Obtain the value of median from the following data of the monthly income of 10 employees of
a company in Rupees.
14,391 15,384 25,591 15,407 16,672 26,522 16,777 26,753 27,850 37,490

Median- Discrete Series


9. From the following data find the value of median:

Income (Rs) 15,000 15,500 16,800 18,000 18,500 17,800


No. of persons 24 26 20 16 6 30

Median- Continuous Series


10. Calculate the median for the following frequency distribution:

Marks 5-10 10-15 15-20 20-25 25-30 30-35 35-40 40-45 45-50
[Link] students 7 15 24 31 42 30 26 15 10

11. Calculate the median for the following data:


Weight (in gm) 410-419 420-429 430-439 440-449 450-459 460-469 470-479
No. of apples 14 20 42 54 45 18 7

12. Calculate the median for the following data


Marks [Link] students Marks [Link] students
Less than 5 29 Less than 30 644
Less than 10 224 Less than 35 650
Less than 15 465 Less than 40 653
Less than 20 582 Less than 45 655
Less than 25 634

13. Compute median from the following data :


Mid value 115 125 135 145 155 165 175 185 195
Frequency 6 25 48 72 116 60 38 22 3

Calculation of missing frequencies


14. An incomplete distribution is given below:
Variable : 0-10 10-20 20-30 30-40 40-50 50-60 60-70
Frequency 10 20 ? 40 ? 25 15

(i) You are given that the median value is 35. Find out missing frequency (given the total
frequency = 170)
(ii) Calculate the arithmetic mean of the completed table.

Mode- Individual Observations


15. Calculate the mode from the following data of the marks obtained by 10 students:
Serial Number Marks obtained Serial Number Marks obtained
1 10 6 27
2 27 7 20
3 24 8 18
4 12 9 15
5 27 10 30

Mode-Discrete Series
16. Calculate the value of mode from the following data:
Marks : 10 15 20 25 30 35 40
Frequency : 8 12 36 35 28 18 9

Mode-Continuous Series
17. Calculate mode from the following data:
Marks [Link] students Marks [Link] students
Above 0 80 Above 60 28
Above 10 77 Above 70 16
Above 20 72 Above 80 10
Above 30 65 Above 90 8
Above 40 55 Above 100 0
Above 50 43

[Link] mode from the following data:


Weight 93-97 98-102 103-107 106-112 113-117 118-122 123-127 128-
132
[Link]
students 2 5 12 17 14 6 3 1

19. From the following data of 122 persons, determine the modal weight:
Weight : 100-110 110-120 120-130 130-140 140-150 150-160 160-170 170-180
[Link]
Persons : 4 6 20 32 33 17 8 2

Standard Deviation - Individual Observations


20. Blood serum cholesterol levels of 10 persons are as under:
240 260 290 245 255 288 272 263 277 251
Calculate standard deviation with the help of assumed mean.

21. Calculate the standard deviation from the following observations:


240.12 240.13 240.15 240.12 240.17
240.15 240.17 240.16 240.22 240.21

Standard deviation- Discrete Series


22. Calculate the standard deviation from the data given below:
Size of the item 3.5 4.5 5.5 6.5 7.5 8.5 9.5
Frequency 3 7 22 60 85 32 8

23. The annual salaries of a group of employees are given in the following table:
Salaries (Rs) 45 50 55 60 65 70 75 80
[Link] persons 3 5 8 7 9 7 4 7

Standard Deviation- Continuous Series


24. Calculate mean, median and Standard Deviation from the following data:
Profits : 0-10 10-20 20-30 30-40 40-50 50-60
[Link] companies 12 17 23 39 16 3

25. Find the standard deviation of the following distribution:


Age 20-25 25-30 30-35 35-40 40-45 45-50
[Link] persons 170 110 80 45 40 35

26. The following are some of the particulars of the distribution of weight of boys and girls in a
class.
Boys Girls
Number 100 50
Mean weight 60 Kg 45 Kg
Variance 9 4
(a) Calculate Standard Deviation of the combined data
(b) Which of the two distributions is more variable?

27. Calculate the mean wages and standard deviation of all the workers taken together.
Section Number of workers Mean wages (in Rs) Standard deviation (in Rs)
Employed
A 50 11130 600
B 60 11200 700
C 90 11150 800

28. The following table shows the monthly expenditure of 80 students of a University on morning
breakfast:
Expenditure [Link] students Expenditure [Link] students
780-820 2 530-570 13
730-770 6 480-520 9
680-720 7 430-470 7
630-670 12 380-420 4
580-620 18 330-370 2

Calculate arithmetic mean, Standard Deviation and Coefficient of variation of the above data.
29. From the prices of shares of X and Y below, find out which is more stable in value:
X 55 54 52 53 56 58 52 50 51 49
Y 108 107 105 105 106 107 104 103 104 101

30. The mean and standard deviation of a set of 100 observations were worked out as 40 and 5
respectively by a computer which by mistake took the value 50 in place of 40 for one
observation. Find the correct mean and variance.

31. The number of employees, wages per employee and the variance of the wages per
employees for two factories are given below.
Factory A Factory B
Number of employees 100 150
Average wage per employee per week (Rs) 3200 2800
Variance of the wages per employee per week (Rs) 625 729
(a) In which factory is there greater variability in the distribution of wages per employee?
(b) Suppose in factory B, the wages of an employee were wrongly noted as Rs3050 instead
of Rs 3650, what would be the correct variance for factory B?

1.2 Problems on Multivariate Summaries

[Link] Karl Pearson’s coefficient of correlation from the following data and interpret its
value:
Roll Number of students : 1 2 3 4 5
Marks in Accountancy : 48 35 17 23 47
Marks in Statistics : 45 20 40 25 45

2. Making use of the data summarised below, calculate the coefficient of correlation
Case X₁ X₂ Case X₁ X₂
A 10 9 E 12 11
B 6 4 F 13 13
C 9 6 G 11 8
D 10 9 H 9 4

3. Calculate the coefficient of correlation from the data given below by using direct method.

X :9 8 7 6 5 4 3 2 1
Y :15 16 14 13 11 12 10 8 9

4. Calculate the coefficient of correlation between X and Y from the following data and calculate
probable error. Assume 69 and 112 as the mean value for X and Yrespectively.

X : 78 89 99 60 59 79 68 61
Y : 125 137 156 112 107 136 123 108
Calculation of Correlation in Grouped Data

5. The following table gives the frequency according to groups of marks obtained by 67 students
in an intelligence test. Measure the degree of relationship between age and intelligence test.

Age in years Total


Test Marks 18 19 20 21
200-250 4 4 2 1 11
250-300 3 5 4 2 14
300-350 2 6 8 5 21
350-400 1 4 6 10 21
Total 10 19 20 18 67

6. The following are the marks obtained by 24 students of a class in Statistics and Accountancy
in an oral examination out of 20 in each subject:

Roll Number Marks in Marks in Roll number Marks in Marks in


of students Statistics Accountancy of students Statistics Accountancy
1 15 13 13 14 11
2 0 1 14 9 3
3 1 2 15 8 5
4 3 7 16 13 4
5 16 8 17 10 10
6 2 9 18 13 11
7 18 12 19 11 14
8 5 9 20 11 7
9 4 17 21 12 18
10 17 16 22 18 15
11 6 6 23 9 15
12 19 18 24 7 3

Prepare a correlation table taking the magnitude of each class interval as four marks and the
first interval as equal to 0 and less than 4. Calculate Karl Pearson’s coefficient of correlation
between the marks in Statistics and marks in accountancy and comment.

1.3 Problems on First Order and Second Order Coefficients

1. On the basis of the following information, compute:


(i) r₂₃.₁ (ii) r₁₃.₂ (iii) r₁₂.₃ where,
r₁₂ = 0.70, r₁₃ = 0.61 and r₂₃ = 0.40

2. On the basis of observations made on 39 cotton plants, the total correlation of yield of
cotton (X₁), the number of bolls i.e., seed vessels (X₂) and height (X₃) are found to be :
r₁₂ = 0.8, r₁₃ = 0.65 and r₂₃ = 0.7

3. If r₁₂ = 0.86, r₁₃ = 0.65 and r₂₃ = 0.72, find the partial correlation coefficient r₁₂.₃
4. Is it possible to get the following from a set of experimental data:
(a) r₂₃ = 0.8, r₁₃ = -0.5, r₁₂ = 0.6
(b) r₂₃ = 0.7, r₁₃ = -0.4, r₁₂ = 0.6

1.4 Problems on Coefficient of Multiple Correlation.

1. The following zero-order correlation coefficients are given


r₁₂ = 0.98, r₁₃ = 0.44 and r₂₃ = 0.54.
Calculate multiple correlation coefficient treating first variable as dependent and second
and third variables as independent.

Common questions

Powered by AI

Multiple correlation coefficients measure the overall strength of association between one dependent variable and multiple independent variables. It is calculated using regression models to assess how well a group of independent variables can predict the dependent variable. The coefficient (R) ranges from 0 to 1, where values closer to 1 indicate a strong linear relationship between the combined predictors and the dependent variable. It provides insights into model fit and predictive power, revealing the collective contribution of multiple factors on an outcome .

To calculate the mode in a continuous frequency distribution, identify the modal class, which has the highest frequency. Apply the mode formula: Mode = L + [(f1 - f0) / (2f1 - f0 - f2)] * h, where L is the lower boundary of the modal class, f1 is the frequency of the modal class, f0 is the frequency of the class preceding the modal class, f2 is the frequency of the class following it, and h is the class width. In discrete distributions, the mode is simply the value with the highest frequency without needing a formula, as exact frequencies are known .

The coefficient of variation (CV) is calculated as the ratio of the standard deviation to the mean, expressed as a percentage. It provides a relative measure of variability in comparison to the mean, making it useful for comparing datasets with different units or scales. Unlike standard deviation, CV is unitless and allows direct comparison regardless of differing magnitudes or units in the datasets, as it indicates the degree of variation relative to the mean .

To correct the arithmetic mean when an error in data entry has occurred, such as reading 83 instead of 53, first determine the sum of the values based on the misread data. Subtract the erroneous entry and add the correct value to get the revised sum. Divide this revised sum by the total number of observations to get the corrected mean. For example, if the mean for 100 students was 40 (sum = 4000) and 83 was read instead of 53, then subtract 83 and add 53 to the total, resulting in a corrected sum of 3970. The corrected mean would be 3970/100 = 39.7 .

To find missing frequencies in a frequency distribution with a known median, first use the cumulative frequency method to determine where the median lies. In a class interval containing the median, use the formula Median = L + [(N/2 - CF) / f] * h, where L is the lower boundary of the median class, N is the total frequency, CF is the cumulative frequency prior to the median class, f is the frequency of the median class, and h is the class width. Rearrange to solve for any unknown frequencies. In the given scenario with a median of 35, adjust the cumulative frequency appropriately to ensure it matches with the given total frequency and given median .

Karl Pearson’s coefficient of correlation is calculated as the covariance of the two variables divided by the product of their standard deviations. This standardized measure ranges from -1 to 1, indicating the strength and direction of a linear relationship between the datasets. A value close to 1 implies a strong positive correlation, meaning as one variable increases, so does the other; a value close to -1 implies a strong negative correlation; and a value close to 0 suggests no linear correlation. This coefficient is crucial in statistical analysis to assess the degree of relationship and predictability between variables .

To determine which distribution shows greater variability, compare their standard deviations or variances. A higher standard deviation or variance indicates greater data spread relative to the mean, therefore more variability. The standard deviation is the square root of the variance, providing a scale that allows for direct comparison across distributions. The choice between using standard deviation or variance can depend on context or ease of interpretation, but in essence, the distribution with the larger measure (either variance or standard deviation) is more variable .

For grouped data, the arithmetic mean is calculated using the mid-point of each class interval multiplied by the class frequency, summing these products, and dividing by the total frequency. This method differs from ungrouped data where each data value is directly used in calculations, as the specific data points within grouped intervals are unknown. Grouped data requires approximation, using class midpoints as representative values assumes uniform distribution within intervals .

When combining data from two distinct groups, the combined standard deviation accounts for the variance within each group as well as the variance between the groups' means. The formula for the combined variance is a weighted sum of the variances of the individual groups, considering the differences in their means. The combined standard deviation is the square root of this combined variance. The combined dataset adds additional variability if the two group means differ significantly. Thus, attention should be paid to relative group sizes and mean differences .

When using partial correlation coefficients, it is important to recognize the influence of a control variable that may obscure the relationship between the other two variables. Partial correlation assesses the relationship between two variables while holding a third constant, isolating the direct correlation separate from confounding influences. Accurate data measurement, understanding variable interactions, and robust statistical techniques are crucial to ensure the control variable’s influence is effectively mitigated and the partial correlation reflects true relationships .

You might also like