0% found this document useful (0 votes)
8 views5 pages

Multivariate Analysis Techniques for HR and Operations Issues

This is the question for practice set of Advance Statistical Method.

Uploaded by

Neha Kushwaha
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
8 views5 pages

Multivariate Analysis Techniques for HR and Operations Issues

This is the question for practice set of Advance Statistical Method.

Uploaded by

Neha Kushwaha
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd

ASM – Question Bank for Practice Only

EC2 pattern – 7 March 2025


Q1. Consider yourself as a manager of production facility. The production facility is located
at Bawana, Delhi. The high temperature of facility, faulty layout, absence of canteen etc.,
makes it very difficult for the workers to work and leads to problems like worker absenteeism
and un-satisfaction. You want to reorganize the facility system so that the working condition
of facility is enhanced and for that you carry out the exploratory research that affects the
working condition. The basic understanding shows that the working condition is unobserved
variable and has to measured on basis of metric variables like temperature etc.
i. Which Multivariate Data Analysis Technique you will use to address the above stated
problem? And Why?
ii. What are the variables that you think helps you to address the problems? (10 marks)
Q-2 a. Multiple regression analysis requires certain assumptions to apply, validate and
interpret the results. Explain the assumptions.
b. How you will deal with the situation in case the assumptions are violated?(5+5 marks)

Q-3 The problem is to understand and predict the tenure of employee for the upcoming
period. Here the variable left is 1 denoting retained, and 0 denoting as left of firm

i. Specify all the assumptions related to the discriminant analysis and comment upon
the applicability of the technique.
ii. Comment upon the purpose of the following Discriminant results.
a. Discriminant Function,
b. Canonical Correlation
c. Model fit
d. Classification Matrix (5+5 Marks)

Q4. Consider yourself as a HR manager of software company MEGAX . The IT services are
provided in the firm and it is located at Bengaluru, India. The HR manager is facing a
problem of workers retention in the firm. The firm is very renowned for the talented
employees. Firm also investing a good amount (money as well as time) in identification of
young talent from colleges and provide them a quality training to make them industry ready.
But once employees spent good amount of time in training and development, they join
competitive firms at higher pay scale and designation. You want to address the problem in
talent retention so that the employee retention is enhanced at MEGA X and for that you carry
out the exploratory research that affects the talent retention. The basic understanding shows
that the talent retention is unobserved variable and has to measured on basis of metric
variables like pay scale, skills etc.

Explain the 6 step approach to deal with the problem w.r.t to the technique used.
Objective, Designing ( variable selection, sample size), Assumptions, Assessment,
Interpretation, Validation. (2+2+6 Marks)

Q-5 a. Multiple regression analysis requires certain assumptions to apply, validate and
interpret the results. Generally these assumptions are verified using pre and post analysis.
What is the meaning of pre and post assumption here? Explain the assumptions with respect
to pre and post methods.
b. How you will deal with the situation in case when Tolerance goes out of permissible
limits? How it is related to the standard error of the variable? (5+5 Marks)

Q-6 How discriminant analysis is same as Multiple Linear Regression and Principal
component Analysis? What is the role of Box M test in discriminant analysis?
i. As per logistic regression and comment upon the following results.
a. Logistic function,
b. Pseudo R square
c. Classification Matrix (5+5 Marks)
Q-1 Consider yourself as an Operations manager of Edutech company BAJURAM. The
education related services are provided in the firm and it is located at Mumbai, India. The
operation manager is facing a problem of unsatisfied students after the pandemic time. The
firm is very renowned and having large market share but during the pandemic the
competition is increased to many fold. Firm bank upon his talented teachers and live
interaction classes. But due to large freelancer employees and huge variety of subject, the
time management is the problem, due to which the doubt classes are not scheduled after the
pre-recorded classes, results in lot of un-satisfaction of parents and students. You want to
address the problem of un-satisfaction so that the student retention is enhanced at
BAJURAM. For that you carry out the exploratory research that affects the un-satisfaction.
The basic understanding shows that the un-satisfaction is unobserved variable and has to
measure on basis of metric variables like % of content covered, Number of hours of live
lecture etc.
i. Which Multivariate Data Analysis Technique you will use to address the above stated
problem? Justify
ii. What are the variables that you think helps you to address the problems?
iii. Explain the 6 step approach to deal with the problem w.r.t to the technique used.
Objective, Designing (variables selection, sample size), Assumptions, Assessment,
Interpretation, Validation. (2+2+6 Marks)

Q-2 a. Multiple regression analysis requires certain assumptions to apply, validate and
interpret the results. Generally these assumptions are verified using pre and post analysis.
What is the meaning of pre and post assumption here? Explain the assumptions with respect
to pre and post methods.
b. How you will deal with the situation in case
i. when Tolerance goes out of permissible limits?
ii. Residual analysis is showing funnel. (5+5 Marks)

Q- The following output is produced, find the missing values and hence comment upon the
following
a. Test of significance of variable
b. Test of significance of model
c. Standard error of the model
d. Confidence interval of beta’s.
e. R square and adjusted R square
SUMMARY OUTPUT

Regression Statistics
Multiple R 0.975062046
R Square
Adjusted R Square 0.947533776
Standard Error
Observations 50

ANOVA
df SS MS F Significance F
Regression 3 7.57E+10 4.52851E-30
Residual 3.92E+09
Total 49 7.96E+10

Coefficients Standard Error t Stat P-value Lower 95% Upper 95%


Intercept 50122.19299 6572.353 1.05738E-09 36892.73332 63351.65
R&D Spend 0.045147 17.84637376 2.63497E-22 0.714838309 0.896592
Administration -0.026815968 -0.52550675 0.601755108 -0.129531575 0.0759
Marketing Spend 0.027228065 0.016451 1.6550773 0.104716819 -0.005886553 0.060343

Q- Comment upon the following residual plot and Normal probability plot

R&D Spend Residual Plot


20000
Residuals

0
0 50000 100000 150000 200000
-20000

-40000
R&D Spend

Normal Probability Plot


250000
200000
150000
Profit

100000
50000
0
0 20 40 60 80 100 120
Sample Percentile
Q- What short of information you can get from such a bi-variate data analysis plot?

Q- Explain the process to deal with the missing value analysis such that:
a. How you determine is this MAR or MCAR
b. How you deal with the MAR situation
c. How you deal with the MCAR situation
d. Consider you need to have the income information in your study but when you collect
data you find that the column has 80% missing values. How you deal with such a
situation?

Common questions

Powered by AI

Key assumptions of discriminant analysis include multivariate normality, homogeneity of variance-covariance matrices, and independence of observations. This technique is applicable to predict employee retention by distinguishing between employees who stay and those who leave based on features like skills and pay scale. If assumptions hold, it provides accurate classification results through discriminant functions .

When data are MAR, methods like multiple imputation or maximum likelihood estimation are appropriate. These approaches account for the pattern of missing data without introducing bias, assuming the reasons for missingness are related to observed data rather than unobserved factors .

Factor Analysis would be a suitable technique because it can identify underlying variables or factors that explain the pattern of correlations within observed variables, such as temperature, faulty layout, and the absence of a canteen, which contribute to worker dissatisfaction. This technique will help in understanding and improving the unobserved variable, i.e., working condition .

The six-step approach includes: 1) Defining the objective, such as improving talent retention; 2) Designing the study with proper variable selection (e.g., pay scale) and adequate sample size; 3) Ensuring assumptions are met, like normal distribution of metrics; 4) Assessing data using exploratory analysis; 5) Interpreting results with statistical significance; 6) Validating the findings to ensure reliability in decision-making .

Discriminant analysis is similar to multiple linear regression in predicting outcomes; however, it focuses on classifying categories rather than continuous outcomes. It shares PCA's ability to reduce dimensionality by transforming original variables but emphasizes distinguishing groups rather than mere variance extraction .

The Box's M test assesses the equality of covariance matrices across groups in discriminant analysis. A significant result indicates heterogeneity, which can affect classification precision, suggesting that corrections or robust methods may be necessary to ensure valid discriminant functions .

Residual plots help identify non-random patterns, suggesting issues like non-linearity or heteroscedasticity. Normal probability plots assess if residuals are normally distributed, crucial for valid inferential statistics. Deviations from a straight line in these plots indicate assumption violations .

To validate the assumptions of multiple regression analysis, pre-analysis checks such as linearity, normality, and homoscedasticity are evaluated using plots and statistical tests. Post-analysis involves checking residuals for randomness and constant variance. If assumptions are violated, it may lead to biased estimates and unreliable predictions. Corrective actions include data transformation, adding interaction terms, or using robust regression techniques .

Multicollinearity inflates the standard errors of the coefficients in a regression model, making them less precise. It reduces statistical power and complicates determining the effect of predictor variables independently, leading to less reliable estimates .

When tolerance is too low, indicating multicollinearity, strategies like removing highly correlated predictors, combining predictors using Principal Component Analysis (PCA), or regularizing methods like Ridge Regression can be employed. It addresses redundancy and stabilizes the coefficients, ensuring a reliable model .

You might also like