0% found this document useful (0 votes)
22 views16 pages

Statistical Concepts and Hypothesis Testing

The document contains a series of multiple-choice questions covering various statistical concepts, including measurement scales, probability, sampling distributions, hypothesis testing, and ANOVA. Each question presents a scenario or concept and provides four answer options. The content is structured in chapters, with questions ranging from basic definitions to more complex statistical theories and calculations.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
22 views16 pages

Statistical Concepts and Hypothesis Testing

The document contains a series of multiple-choice questions covering various statistical concepts, including measurement scales, probability, sampling distributions, hypothesis testing, and ANOVA. Each question presents a scenario or concept and provides four answer options. The content is structured in chapters, with questions ranging from basic definitions to more complex statistical theories and calculations.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

CHAPTER 1-7

1. The scale of measurement that is used to rank order the observation for a variable is
called the
a. ratio scale.
b. ordinal scale.
c. nominal scale.
d. interval scale.

2. Five hundred residents of a city are polled to obtain information on voting intentions in
an upcoming city election. The five hundred residents in this study is an example of a(n)
a. census.
b. sample.
c. observation.
d. Population.

3. What is the median of 26, 30, 24, 32, 32, 31, 27 and 29?
A. 32
B. 29
C. 30
D. 29.5

4. The weights (in grams) of the contents of several small bottles are 4, 2, 5, 4, 5, 2 and 6.
What is the sample variance?
A. 6.92
B. 4.80
C. 1.96
D. 2.33

5. Chebyshev’s Theorem:
A applies to all samples.
B applies only to samples from a normal population.
C gives a narrower range of predictions than the Empirical Rule.
D is based on Sturges’ Rule for data classification.
The strength of Chebyshev’s Theorem is that it makes no assumption about normality,
while the E.R. only works for normal populations.

6. Events A and B are mutually exclusive when:


A. their joint probability is zero.
B. they are independent events.
C. P(A)P(B) = 0
D. P(A)P(B) = P(A | B)

7. Independent events A and B would be consistent with which of the following


statements:
A) P ( A ) = .3, P ( B ) = .5, P ( A u B ) = .4.
B) P ( A ) = .4, P ( B ) = .3, P ( A u B ) = .5.
C) P ( A ) = .4, P ( B ) = .5, P ( A u B ) = .2.
D) P ( A ) = .5, P ( B ) = .4, P ( A u B ) = .3
By the definition of independent events, P(A)*P(B) = P(A u B)

8. Given the contingency table shown here, find P(V | W).

A. .4000
B. .0950
C. .2375
D. .5875
This is a conditional probability P(V|W) = 19/80.

9. The discrete random variable X is the number of students that show up for Professor
Smith's office hours on Monday afternoons. The table below shows the probability
distribution for X. What is the probability that fewer than 2 students come to office hours
on any given Monday?

A. .10
B. .40
C. .70
D. .90

CHAPTER 7

10. If the random variable Z has a standard normal distribution, then P(Z ≤ -1.72) is:
A. 0.9573.
B. 0.0446.
C. 0.5016.
D. 0.0427.

CHAPTER 8
11. A sampling distribution describes the distribution of:
A. a parameter.
B. a statistic.
C. either a parameter or a statistic.

12. neither a parameter nor a statistic.


Sampling error can be avoided:
A. by using an unbiased estimator.
B. by eliminating nonresponses (e.g., older people).
C. by no method under the statistician's control.
D. either by using an unbiased estimator or by eliminating nonresponse.

13. The Central Limit Theorem (CLT) implies that:


A. the population will be approximately normal if n ≥ 30.
B. repeated samples must be taken to obtain normality.
C. the distribution of the mean is approximately normal for large n.
D. the mean follows the same distribution as the population.

14. The owner of Limp Pines Resort wanted to know the average age of its clients. A
random sample of 25 tourists is taken. It shows a mean age of 46 years with a standard
deviation of 5 years. The width of a 98 percent CI for the true mean client age is
approximately:
A. ± 1.711 years.
B. ± 2.326 years.
C. ± 2.492 years.
D. ± 2.797 years.

15. To estimate the average annual expenses of students on books and class materials a
sample of size 36 is taken. The sample mean is $850 and the sample standard deviation is
$54. A 99 percent confidence interval for the population mean is:
A. $823.72 to $876.28
B. $832.36 to $867.64
C. $826.82 to $873.18
D. $825.48 to $874.52

16. The standard error of the mean decreases when:


A. the sample size decreases.
B. the standard deviation increases.
C. the standard deviation decreases or n increases.
D. the population size decreases.

17. A poll showed that 48 out of 120 randomly chosen graduates of California medical
schools last year intended to specialize in family practice. What is the width of a 90
percent confidence interval for the proportion that plan to specialize in family practice?
A. ± .0447
B. ± .0736
C. ± .0876
D. ± .0894

18. A financial institution wishes to estimate the mean balances owed by its credit card
customers. The population standard deviation is estimated to be $300. If a 99 percent
confidence interval is used and an interval of ± $75 is desired, how many cardholders
should be sampled?
A) 3382
B) 629
C) 87
D) 107
19. Which statement is incorrect? Explain.
A. If p = .50 and n = 100, the standard error of the sample proportion is .05.
B. In a sample size calculation for estimating π, it is conservative to assume π = .50.
C. If n = 250 and p = .06, we cannot assume normality in a confidence interval for
π.

19. In point estimation


A. data from the population is used to estimate the population parameter
B. data from the sample is used to estimate the population parameter
C. data from the sample is used to estimate the sample statistic
D. the mean of the population equals the mean of the sample

20. A population has a mean of 75 and a standard deviation of 8. A random sample of


800 is selected. The expected value of x is
a. 8
b. 75
c. 800
d. None of these alternatives is correct

CHAPTER 9

21. After testing a hypothesis, we decided to reject the null hypothesis. Thus, we are
exposed to:
A. Type I error.
B. Type II error.
C. Either Type I or Type II error.
D. Neither Type I nor Type II error.

22. Which of the following is incorrect?


A. The level of significance is the probability of making a Type I error.
B. Lowering both α and β at once will require a higher sample size.
C. The probability of rejecting a true null hypothesis increases as n increases.
D. When Type I error increases, Type II error must decrease, ceteris paribus
23. Which statement is correct about a p-value?
A. The smaller the p-value the stronger the evidence in favor of the alternative
hypothesis.
B. The smaller the p-value the stronger the evidence in favor the null hypothesis
C. Whether a small p-value provides evidence in favor of the null hypothesis depends
on whether the test is one-sided or two-sided.
D. Whether a small p-value provides evidence in favor of the alternative hypothesis
depends on whether the test is one-sided or two-sided.

24. "Currently, only 20% of arrested drug pushers are convicted," cried candidate
Courageous Calvin in a campaign speech. "Elect me and you'll see a big increase in
convictions." A year after his election a random sample of 144 case files of arrested drug
pushers showed 36 convictions. For a right-tailed test, the p-value is approximately
a. 0.0435
b. 0.9332
c. 0.0250
d. 0.0668

25. A meteorologist stated that the average temperature during July in Chattanooga was
80 degrees. A sample of July temperatures over a 32-year period was taken. The correct
set of hypotheses is .
a. H0: μ < 80 Ha: μ ≤ 80
b. H0: μ ≤ 80 Ha: μ > 80
c. H0: μ ≠ 80 Ha: μ = 80
d. H0: μ = 80 Ha: μ ≠ 80

26. For a two-tailed test with a sample size of 40, the null hypothesis will NOT be
rejected at a 5% level of significance if the test statistic is .
a. between -1.96 and 1.96, exclusively
b. greater than 1.96
c. less than 1.645
d. greater than -1.645

27. A sample of 16 ATM transactions shows a mean transaction time of 67 seconds


with a standard deviation of 12 seconds. Find the test statistic to decide whether
the mean transaction time exceeds 60 seconds.
A. 1.457
B. 2.037
C. 2.333
D. 1.848

28. Last year, 10 percent of all teenagers purchased a new iPhone. This year, a
sample of 260 randomly chosen teenagers showed that 39 had purchased a new
iPhone. To test whether the percent has risen, the p-value is approximately:
A. .0501
B. .0314
C. .0492
D. .0036

29. A sample of 16 ATM transactions shows a mean transaction time of 67 seconds


with a standard deviation of 12 seconds. Find the critical value to test whether the
mean transaction time exceeds 60 seconds at α = .01.
A. 2.947
B. 2.602
C. 2.583
D. 2.333

30. Ajax Peanut Butter's quality control allows 2 percent of the jars to exceed the
quality standard for insect fragments. A sample of 150 jars from the current day's
production reveals that 30 exceed the quality standard for insect fragments. Which is
incorrect?
A. Normality of p may safely be assumed in the hypothesis test.
B. A right-tailed test would be appropriate.
C. Common sense suggests that quality control standards aren't met.
D. Type II error is more of a concern in this case than Type I error.

CHAPTER 10

31. Which of the following is an example of a two-sample hypothesis test?


a. Is the average service time at a restaurant different than 20 minutes?
b. Is the average service time at a restaurant different between Friday and
Saturday night?
c. Is average service time at a restaurant less than 20 minutes?
d. Is the average service time at a restaurant more than 20 minutes?

32. Does the Speedo Fastskin II Male Hi-Neck Bodyskin competition racing swimsuit
improve a swimmer's 200-yard individual medley performance times? A test of 100
randomly chosen male varsity swimmers at several different universities showed that 66
enjoyed improved times, compared with only 54 of 100 female varsity swimmers. To test
for equality in the proportions of men versus women who experienced improvement, the
test statistic is approximate:
A. 1.73
B. 1.47
C. 2.31
D. Can't tell without knowing the tail of the test.

33. We pool the sample variances when


a. Comparing two paired means
b. Comparing two means with unknown variances (assumed unequal)
c. Comparing two means with known variances
d. Comparing two means with unknown variances (assumed equal)
34. Carver Memorial Hospital's surgeons have a new procedure that they think will
decrease the time to perform an appendectomy. A sample of 8 appendectomies using the
old method had a mean of 38 minutes with a variance of 36 minutes, while a sample of
10 appendectomies using the experimental method had a mean of 29 minutes with a
variance of 16 minutes. For a right-tail test of means (assume equal variances) the pooled
variance is
A. 14.76
B. 26.00
C. 24.75
D. 27.54

35. In a left-ailed test comparing two means with variances unknown but assumed to be
equal, the sample sizes were n1 = 8 and n2 = 12. At α = .05, the critical value would be:
A. -1.960
B. -2.101
C. -1.734
D. -1.645

36. Of 200 youthful gamers (under 18) who tried the new Z-Box-Plus game, 160 rated it
"excellent," compared with only 144 of 200 adult gamers (18 or over). The 95 percent
confidence interval for the difference of proportions would be approximately:
A. [+.013,+.263].
B. [-.014, +.188].
C. [-.003, +.163].
D. [+.057,+.261].

37. During a test period, an experimental group of 10 vehicles using an 85%


ethanol-gasoline mixture showed mean CO2 emissions of 240 pounds per 100 miles, with
a standard deviation of 20 pounds. A control group of 14 vehicles using regular gasoline
showed mean CO2 emissions of 252 pounds per 100 miles with a standard deviation of
15 pounds. At a = 0.05, in a left-tailed test, the critical value to compare the means
(assuming equal variances) is
A. -2.508
B. -2.074
C. -1.321
D. -1.717

38. Two well-known aviation training schools are being compared using random samples
of their graduates. It is found that 70 of 140 graduates of Fly-More Academy passed their
FAA exams on the first try, compared with 104 of 260 graduates of Blue Yonder Institute.
To compare the pass rates, the pooled proportion would be:
A. .500
B. .435
C. .400
D. .345

39. The table below shows the mean number of daily errors by air traffic controller
trainees during the first two weeks on the job. We want to perform a paired t-test at α =
.05 to see if the mean daily errors decreased significantly.

The test statistic is


A. 1.25
B. 1.75
C. 0.87
D. 0.79

40. Group 1 has a mean of 13.4 and group 2 has a mean of 15.2. Both populations are
known to have a variance of 9.0 and each sample consists of 18 items. What is the test
statistic to testfor equality of population means?
A. -1.755
B. -1.643
C. -1.800
D. -1.285
With known variances, zcalc = (13.4 - 15.2)/[9.0/18 + 9.0/18]1/2 = -1.800
CHAPTER 11

41. Variation "within" the ANOVA treatments represents:


A. random variation.
B. differences between group means.
C. differences between group variances.
D. the effect of sample size.

42. In an ANOVA, when would the F-test statistic be zero?


A. When there is no difference in the variances.
B. When the treatment means are the same.
C. When the observations are normally distributed.
D. The F-test statistic cannot ever be zero.

43. Degrees of freedom for the between-group variation in a one-factor


ANOVA with n1 = 8, n2 = 5, n3 = 7, n4 = 9 would be:
A. 28
B. 3
C. 29
D. 4

44. Given the following ANOVA table (some information is missing), find
the critical value of F.05

A. 3.06
B. 2.90
C. 2.36
D. 3.41

45. For this one-factor ANOVA (some information is missing), how many treatment
groups were there?

A. Cannot be determined
B. 3
C. 4
D. 2

46. Refer to the following partial ANOVA results from Excel (some
information is missing).

Degrees of freedom for between groups variation are:


A. 3
B. 4
C. 5
D. Can't tell from given information.

47. The Internal Revenue Service wishes to study the time required to process tax returns
in three regional centers. A random sample of three tax returns is chosen from each of
three centers. The time (in days) required to process each return is recorded as shown
below.
The test to use to compare the means for all three groups would require:
A. three-factor ANOVA.
B. one-factor ANOVA.
C. repeated two-sample test of means.
D. two-factor ANOVA with replication.

48. To compare the cost of three shipping methods, a random sample of four shipments is
taken for each of three firms. The cost per shipment is shown below.

In a one-factor ANOVA, degrees of freedom for the within-groups sum


of squares will be:
A. 11
B. 3
C. 9
D. 2

49. Here is an Excel ANOVA table for an experiment that analyzed two factors that may
affect patients' blood pressure (some information is missing).
At α = .10 the interaction is:
A. significant.
B. insignificant.
C. borderline.

50. Which Excel function gives the right-tail p-value for an ANOVA test with a
test statistic Fcalc = 4.52, n = 29 observations, and c = 4 groups?
A. =[Link](4.52, 3, 25)
B. =[Link](4.52, 4, 28)
C. =[Link](4.52, 4, 28)
D. =[Link](4.52, 3, 25)

CHAPTER 12

51. Which of the following statements regarding the coefficient of correlation is true?
A: It ranges from -1.0 to +1.0 inclusive
B: It measures the strength of the relationship between two variables
C: A value of 0.00 indicates two variables are not related
D: All of the above

52. The variable used to predict another variable is called the:


A. response variable.
B. regression variable.
C. independent variable.
D. dependent variable.

53. Suppose the least squares regression equation is Y' = 1202 + 1,133X. When X = 3,
what does Y' equal?
A: 5,734
B: 8,000
C: 4,601
D: 4,050

54. A hypothesis test is conducted at the 5 percent level of significance to


test whether the population correlation is zero. If the sample consists of
25 observations and the correlation coefficient is 0.60, then the
computed test statistic would be:
A. 2.071.
B. 1.960.
C. 3.597.
D. 1.645.

55. If the attendance at a baseball game is to be predicted by the equation Attendance =


16,500 - 75 Temperature, what would be the predicted attendance if Temperature is 90
degrees?
A. 6,750
B. 9,750
C. 12,250
D. 10,020

56. A local trucking company fitted a regression to relate the travel time
(days) of its shipments as a function of the distance traveled (miles). The
fitted regression is Time = -7.126 + .0214 Distance, based on a sample
of 20 shipments. The estimated standard error of the slope is 0.0053.
Find the critical value for a right-tailed test to see if the slope is positive,
using α = .05.
A. 2.101
B. 2.552
C. 1.960
D. 1.734

57. Based on the regression equation, we can


A: predict the value of the dependent variable given a value of the independent variable
B: predict the value of the independent variable given a value of the dependent variable
C: measure the association between two variables
D: all of the above

58. In the least squares equation, Y' = 10 + 20X the value of 20 indicates
A: the Y intercept
B: for each unit increase in X, Y increases by 20.
C: for each unit increase in Y, X increases by 20
D: none of the above.

59. Which of the following is not a characteristic of the F-test in a simple


regression?
A. It is a test for overall fit of the model.
B. The test statistic can never be negative.
C. It requires a table with numerator and denominator degrees of freedom.
D. The F-test gives a different p-value than the t-test.

60. Amelia used a random sample of 100 accounts receivable to estimate the relationship
between Days (number of days from billing to receipt of payment) and Size (size of
balance due in dollars). Her estimated regression equation was Days = 22 + 0.0047 Size
with a correlation coefficient of .300. From this information we can conclude that:
A. 9 percent of the variation in Days is explained by Size.
B. autocorrelation is likely to be a problem.
C. the relationship between Days and Size is significant.
D. larger accounts usually take less time to pay.

Common questions

Powered by AI

The Central Limit Theorem implies that the distribution of the sample mean approaches a normal distribution as the sample size becomes large, regardless of the population's distribution. This fundamental theorem allows statisticians to make inferences about population parameters using sample data, facilitating hypothesis testing and confidence interval estimation even when the population distribution is unknown .

Type I errors occur when a true null hypothesis is incorrectly rejected, while Type II errors happen when a false null hypothesis is not rejected. Increasing the sample size can mitigate these risks by providing more accurate estimates of the population parameters, thereby reducing the probability of both errors. With larger samples, the test results become more reliable, increasing the power of the test and allowing for better discrimination between true and false hypotheses .

Pooling sample variances is significant because it combines the variances of two samples, leading to a more accurate and efficient estimate of the overall variance when comparing means. This method assumes that the populations from which the samples are drawn have equal variances (homogeneity of variance), which allows for a valid computation of a pooled variance when conducting t-tests for comparing means .

Chebyshev’s Theorem applies to all samples without the requirement of a normal distribution, making no assumptions about normality. Unlike the Empirical Rule, which provides predictions solely for normal distributions, Chebyshev’s Theorem gives a broader application as it can be used for any distribution shape .

The sample variance provides an understanding of variability by measuring how spread out the numbers in a data set are around the mean. For the given set of weights (4, 2, 5, 4, 5, 2, and 6), the sample variance is calculated as 4.80 grams squared, reflecting the average of the squared differences from the Mean .

The ordinal scale is used to rank order observations for a variable because it categorizes and orders the data according to the natural order of the categories . This makes it suitable for arranging data in a sequence, though without expressing the magnitude of difference between them.

Regression analysis predicts a response variable based on the relationship established by an independent variable, using a linear equation defined by known data points. The coefficient of correlation provides insight into the strength and direction of this relationship; its value ranges from -1 to +1, with values close to 1 or -1 indicating a strong relationship and a value close to 0 suggesting a weak relationship. This allows researchers to identify how changes in the predictor might be associated with variation in the response variable .

A sampling distribution describes the distribution of a sample statistic (like the mean) calculated from a sample of a population. It's crucial because it forms the foundation for making inferences about the population based on sample data, allowing statisticians to estimate population parameters and measure the reliability of these estimates using measures like standard error .

Understanding mutually exclusive events is crucial because it affects how probabilities are calculated, as mutually exclusive events cannot occur simultaneously. Two events are mutually exclusive if their joint probability is zero, indicating that the occurrence of one event excludes the occurrence of the other .

The F-test in ANOVA examines whether there are significant differences between group means by comparing the variance between groups to the variance within groups. The degrees of freedom determine the critical value of the F-test based on the number of groups and the total sample size, dictating the thresholds for statistical significance. In a one-factor ANOVA, the degrees of freedom between groups are based on the number of groups minus one, while within groups are based on the total number of observations minus the number of groups .

You might also like