0% found this document useful (0 votes)
14 views9 pages

Statistical Tests and Hypothesis Testing

The document outlines various statistical concepts including types of variables, hypotheses, confidence intervals, and how to select appropriate statistical tests. It provides examples of hypothesis testing using t-tests, ANOVA, regression, and Chi-square tests, along with interpretations of results. Good practices for statistical testing are also emphasized, such as checking assumptions and understanding the significance of results.

Uploaded by

anwar zouhri
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
14 views9 pages

Statistical Tests and Hypothesis Testing

The document outlines various statistical concepts including types of variables, hypotheses, confidence intervals, and how to select appropriate statistical tests. It provides examples of hypothesis testing using t-tests, ANOVA, regression, and Chi-square tests, along with interpretations of results. Good practices for statistical testing are also emphasized, such as checking assumptions and understanding the significance of results.

Uploaded by

anwar zouhri
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

‭1.

Types of Variables‬

‭●‬ ‭Quantitative (Scale):‬‭numeric values (e.g., scores,‬‭age, salary)‬

‭●‬ ‭Categorical (Nominal/Ordinal):‬‭groups or categories‬‭(e.g., gender, department)‬

‭Understanding your variable type is‬‭essential‬‭to selecting the correct test.‬

‭2. Types of Hypotheses‬

‭●‬ ‭Null Hypothesis (H₀):‬‭No difference or no effect‬

‭●‬ ‭Alternative Hypothesis (H₁):‬‭There is a difference or effect‬

‭●‬ ‭Hypothesis testing relies on‬‭p-values‬‭to decide whether to reject H₀‬

‭3. Confidence Intervals‬

‭●‬ A
‭ ‬‭95% confidence interval‬‭gives a range in which we’re‬‭95% confident the‬‭true‬
‭population parameter‬‭lies‬

‭●‬ ‭Used to support hypothesis test results‬

‭4. How to Choose the Right Statistical Test‬

‭You choose the test based on:‬

‭●‬ ‭The‬‭type of variables‬‭(categorical vs. scale)‬

‭●‬ ‭The‬‭number of groups‬‭or predictors‬

‭●‬ ‭Whether you're testing for a‬‭difference‬‭, an‬‭association‬‭,‬‭or a‬‭prediction‬

‭Includes:‬

‭●‬ ‭T-tests‬

‭●‬ ‭ANOVA‬

‭●‬ ‭Regression (simple, multiple, logistic)‬


‭●‬ ‭Chi-square‬

‭5. Good Practices & Warnings‬

‭●‬ ‭Always‬‭check assumptions‬‭: normality, independence,‬‭sample size‬

‭●‬ ‭Don't run tests blindly — first explore the data (descriptive stats, graphs)‬

‭●‬ ‭Understand what‬‭each test result actually means‬‭(statistical‬‭vs. practical significance)‬

‭ e notice here that the‬‭average Test_Score‬‭in the population is significantly different from the‬
W
‭reference value of 50.‬

‭To validate this using a‬‭One-Sample T-Test‬‭, we test‬‭the following hypotheses:‬

‭●‬ H
‭ ₀‬‭: There is no significant difference between the‬‭population’s average score and the‬
‭reference value of 50.‬

‭●‬ H
‭ ₁‬‭: There is a significant difference between the‬‭population’s average score and the‬
‭reference value of 50.‬

‭ ith a‬‭t-value of 56.543‬‭,‬‭p < .001‬‭, and a‬‭mean difference‬‭of 26.09‬‭, we clearly reject the null‬
W
‭hypothesis.‬
‭This means that the average Test_Score (M = 76.09) is‬‭statistically and substantially higher‬
‭than 50.‬

‭ o based on the result of our t-test we can say the valid hypothese is‬‭H1‬‭because the p value is <‬
S
‭0.05‬

‭ e notice here that there’s‬‭no significant difference‬‭between male and female participants in‬
W
‭terms of their average test scores (M = 75.98 for males, M = 76.19 for females).‬

‭ o confirm this result using an‬‭Independent Samples‬‭T-Test‬‭, we evaluate the following‬


T
‭hypotheses:‬
‭●‬ H
‭ ₀‬‭: There is no significant difference in the mean Test_Score between male and female‬
‭participants.‬

‭●‬ H
‭ ₁‬‭: There is a significant difference in the mean‬‭Test_Score between male and female‬
‭participants.‬

‭Based on the results:‬

‭●‬ T
‭ he‬‭Levene’s Test‬‭for equality of variances has a p-value of‬‭.967‬‭, so we assume‬‭equal‬
‭variances‬‭.‬

‭●‬ ‭The‬‭t-test p-value‬‭is‬‭.816‬‭, which is‬‭greater than 0.05‬‭.‬

‭ herefore, we‬‭fail to reject the null hypothesis‬‭(H₀), and conclude that‬‭there is no statistically‬
T
‭significant difference‬‭between male and female test scores in this sample.‬

‭ e observe that the mean‬‭Pre_Training_Score‬‭is‬‭60.84‬‭,‬‭while the‬‭Post_Training_Score‬‭is‬


W
‭65.89‬‭.‬
‭Based on this initial data, we can say that training‬‭might have had‬‭an effect on scores.‬

‭To confirm this result, we apply a‬‭paired-samples‬‭t-test‬‭and test the following hypotheses:‬

‭●‬ ‭H₀‬‭: Training had‬‭no effect‬‭on scores.‬

‭●‬ H ‭ ₁‬‭: Training had‬‭an effect‬‭on scores.‬


‭res‬
‭●‬ ‭The‬‭mean difference‬‭is‬‭-5.05‬

‭●‬ ‭The‬‭p-value‬‭is‬‭.000‬‭, which is‬‭less than 0.05‬

‭ ased on the p-value (< 0.05), we‬‭reject H₀‬‭and conclude that‬‭training had a‬
B
‭statistically significant effect on scores‬‭.‬

‭ sing‬‭
U Satisfaction_Score‬ ‭, we compared the mean satisfaction‬‭levels across‬
‭Department groups (HR, IT, Marketing) using a One-Way ANOVA with Tukey post hoc‬
‭test.‬
‭The post hoc results show that:‬

‭●‬ ‭HR has the highest satisfaction score‬

‭●‬ ‭IT is in the middle‬

‭●‬ ‭Marketing has the lowest‬

‭ ll pairwise comparisons are‬‭statistically significant‬‭(p < .001), as indicated by the‬‭Sig. column‬


A
‭in the Tukey test.‬

‭ here are‬‭significant differences‬‭in satisfaction levels between‬‭all three‬


T
‭departments‬‭. Each department’s average satisfaction score is‬‭significantly‬
‭different‬‭from the others.‬

‭ e used a binary logistic regression to predict the likelihood of being hired based on three‬
W
‭variables:‬‭GPA‬‭,‬‭Work_Experience‬‭, and‬‭Interview_Score‬‭.‬

‭ he overall model is‬‭statistically significant‬‭(Chi²‬‭= 15.717, p = .001), which means that at least‬
T
‭one predictor has an effect. The model explains about‬‭17.9%‬‭of the variation in hiring‬
‭(Nagelkerke R² = .179).‬

‭Looking at the individual predictors:‬

‭●‬ I‭ nterview_Score‬‭is statistically significant (‬‭p =‬‭.001‬‭), so we can say it has a‬‭real effect‬
‭on the chance of being hired. The higher the score, the higher the odds (Exp(B) = 1.13).‬

‭●‬ ‭Work_Experience‬‭is‬‭borderline‬‭(p = .053), so it might‬‭have a small effect.‬

‭●‬ ‭GPA‬‭is‬‭not significant‬‭(p = .965), so it doesn't influence hiring in this model.‬

‭Conclusion:‬‭Interview performance is the best predictor‬‭of being hired in this case.‬

‭We used a‬‭simple linear regression‬‭to predict‬‭Salary‬‭based on‬‭Years_Experience‬‭.‬


‭ he model is‬‭statistically significant‬‭(‬‭p = .000‬‭) and explains a‬‭very large portion of the‬
T
‭variation‬‭in salary (‬‭R² = .870‬‭). This means that‬‭87%‬‭of the variation in salary‬‭can be‬
‭explained by years of experience.‬

‭Looking at the coefficients:‬

‭●‬ ‭The‬‭p-value for Years_Experience is .000‬‭, so it's‬‭statistically significant‬‭.‬

‭●‬ T
‭ he‬‭B coefficient is 4103.86‬‭, which means that‬‭each additional year of experience‬
‭increases salary by approximately 4104 units‬‭.‬

‭Conclusion‬‭:‬

‭ ears of experience has a strong and significant impact on salary.‬‭The more‬


Y
‭experience someone has, the higher their salary is likely to be.‬

‭ e tested the association between‬‭Smoking_Status‬‭and‬‭Exercise_Level‬‭using the‬‭Chi-Square‬


W
‭Test of Independence‬‭.‬

‭●‬ ‭H₀‬‭: There is no association between smoking and exercise‬‭level.‬

‭●‬ ‭H₁‬‭: There is an association between smoking and exercise‬‭level.‬

‭Looking at the table:‬

‭●‬ ‭The‬‭Chi-square value = 50.347‬‭, with‬‭df = 2‬

‭●‬ ‭The‬‭p-value = .000‬‭→ which is‬‭less than 0.05‬

‭Conclusion‬‭:‬

‭ e reject H₀. There is a‬‭statistically significant association‬‭between‬


W
‭Smoking_Status‬‭and‬‭Exercise_Level‬‭.‬
‭In other words,‬‭exercise habits differ depending on whether someone smokes or‬
‭not‬‭.‬
‭ 95% confidence interval means that if we repeated the study many times, about 95% of‬
A
‭the calculated intervals would contain the true population parameter.‬

‭ e use a t-test when comparing the means of‬‭two groups‬‭, and a one-way ANOVA when‬
W
‭comparing the means across‬‭three or more groups‬‭.‬

‭One way anova.‬

‭ o compare the means between two groups, you should use a t-test (specifically, an‬
T
‭independent samples t-test).‬

‭ redictor‬
P ‭ utcome‬
O ‭Test to Use‬
‭(Independent)‬ ‭(Dependent)‬

‭Quantitative‬ ‭Quantitative‬ ‭ imple or Multiple Linear‬


S
‭Regression‬

‭Quantitative‬ ‭Categorical‬ ‭Logistic Regression‬

‭Categorical‬ ‭Quantitative‬ ‭ -Test‬‭(if 2 groups) or‬‭ANOVA‬‭(if‬


T
‭3+ groups)‬

‭Categorical‬ ‭Categorical‬ ‭Chi-Square Test (χ²)‬

‭Chi-Square (χ²) Test — When and How (in SPSS)‬

‭Use when:‬

‭●‬ ‭Both variables are‬‭categorical‬

‭●‬ ‭You want to test‬‭association‬‭or‬‭independence‬


‭SPSS Steps:‬

‭pgsql‬

‭CopierModifier‬

Analyze → Descriptive Statistics → Crosstabs → Statistics‬



→ Chi-square‬

H1 < 0.05- there’s a difference‬



H0> there’s no difference‬

On sample t-test= comparing a number‬


Independent t-test eg: gender‬


Paired t-test = same sample before and after‬


#‬
‭ Question‬
‭ Test to Use‬
‭ Reason‬

1‬ ‭
‭ Using the variable Stress_Score,‬ One-Sample‬
‭ You're‬

test whether the average is‬
‭ T-Test‬
‭ comparing a‬

significantly different from 50.‬
‭ sample mean to‬

a fixed value‬

(50)‬

2‬ ‭
‭ Using Test_Score as the outcome,‬ Independent‬
‭ Two independent‬

compare the means between Male‬
‭ Samples‬
‭ groups (Male vs‬

and Female participants.‬
‭ T-Test‬
‭ Female),‬

quantitative‬

outcome‬

3‬ ‭
‭ Using Pre_Training_Score and‬ Paired‬
‭ Same‬

Post_Training_Score, evaluate‬
‭ Samples‬
‭ participants‬

whether training had an effect‬
‭ T-Test‬
‭ tested before‬

on scores.‬
‭ and after‬

training‬

4‬ ‭
‭ Use Satisfaction_Score to‬ One-Way‬
‭ Comparing means‬

compare mean satisfaction levels‬
‭ ANOVA‬
‭ across 3+‬

across Department groups (HR,‬
‭ groups‬

IT, Marketing).‬

5‬ ‭
‭ Model the likelihood of being‬ Binary‬
‭ Outcome is‬

Hired using predictors GPA,‬
‭ Logistic‬
‭ binary (Hired:‬

Work_Experience, and‬
‭ Regression‬
‭ Yes/No),‬

Interview_Score.‬
‭ predictors are‬

quantitative‬

6‬ ‭
‭ Using Years_Experience to‬ Simple‬
‭ One‬

predict Salary.‬
‭ Linear‬
‭ quantitative‬

Regression‬
‭ predictor and‬

one‬

quantitative‬

outcome‬

7‬ ‭
‭ Test whether there is an‬ Chi-Square‬
‭ Both variables‬

association between‬
‭ Test of‬
‭ are categorical‬

Smoking_Status and‬
‭ Independence‬
‭ – checking‬

Exercise_Level.‬
‭ (χ²)‬
‭ association‬

You might also like