0% found this document useful (0 votes)
10 views6 pages

Statistics Project

The document contains a series of questions and answers related to statistics, probability, and data analysis, covering key concepts such as sampling methods, types of data, and statistical measures. It includes exercises on frequency tables, data interpretation, and the application of statistical principles in research scenarios. The content is structured to assess understanding of statistical terminology and concepts.

Uploaded by

Judy Ann Ventura
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
10 views6 pages

Statistics Project

The document contains a series of questions and answers related to statistics, probability, and data analysis, covering key concepts such as sampling methods, types of data, and statistical measures. It includes exercises on frequency tables, data interpretation, and the application of statistical principles in research scenarios. The content is structured to assess understanding of statistical terminology and concepts.

Uploaded by

Judy Ann Ventura
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as DOCX, PDF, TXT or read online on Scribd

CRITICA

KNOWLED COMPREHENSI L TOTA PERCENTA


TOPICS
GE ON THINKIN L GE
G
DEFINITION 1 4,5 12, 13, 8 26.67%
OF 14,15,16
STATISTICS,
PROBABILITY
, AND KEY
TERMS
DATA, 17,18, 14 46.67%
SAMPLING, 19,20,21,2
AND 3 6,7,8,9 2,
VARIATION 23,24,25
IN DATA
SAMPLING
FREQUENCY, 2 26,27,28,2 8 26.67%
FREQUENCY 10, 11 9,30
TABLES, AND
LEVELS OF
MEASUREME
NT
TOTAL 3 8 19 30 100%

1. The branch of science that deals with the collection, analysis, interpretation, and
presentation of data.
A. Probability B. Statictics C. Data Analysis D. Descriptive Statistics

2. The way a set of data is measured is called its _________.

A. Variation B. Levels of Measurement C. Frequency D. Data Value

3. These data are the result of categorizing or describing attributes of a population and are
also often called categorical data.

A. . Quantitative data [Link] Data C. Sampling Data D. Variation

4. . Below is a two-way table showing the types of college sports played by men and women.
Given the data below, calculate the marginal distributions of college sports for the people
surveyed
5. Below is a two-way table showing the types of college sports played by men and women.
Given these data, calculate the conditional distributions for the subpopulation of women who
play college sports.

***Use the following information to answer the next four exercises:***

A study was done to determine the age, number of times per week, and the duration (amount
of time) of residents using a local park in San Antonio, Texas. The first house in the
neighborhood around the park was selected randomly, and then the resident of every eighth
house in the neighborhood around the park was interviewed.

6. The sampling method was


a. simple random b. systematic c. stratified d. cluster

7. Duration (amount of time) is what type of data?


a. qualitative (categorical) b. quantitative discrete c. quantitative continuous

8. . The colors of the houses around the park are what kind of data?
a. qualitative (categorical; b. quantitative discrete c. quantitative continuous

9. The population is ________.

10. What is the frequency table used for?


A. To calculate the mean of the dataset
B. To display the distribution of data values in a tabular format
C. To represent data using a scatter plot
D. To determine correlation between variables

11. In the data set of exam scores, the score 85 appears 6 times out of 50. What is the relative
frequency of 85?

A. 6 B. 0.12 C. 12% D. Both B and C

12. Which of the following best describes the field of statistics?


A. The study of numbers in the abstract.
B. The collection, analysis, interpretation, presentation, and organization of data.
C. The manipulation of numerical values to create theoretical models.
D. The process of making predictions without using data.

13. Which of the following terms refers to the middle value in a sorted dataset?
A. Mode
B. Mean
C. Median
D. Range

14. If a dataset has a mean of 10 and a standard deviation of 2, what is the z-score of a data
point with a value of 12?
A. 0
B. 1
C. -1
D. 2

Answer: B. 1 (Z-score = (X - mean) / standard deviation = (12 - 10) / 2 = 1)

15. A researcher is trying to estimate the average age of people who use a certain
website. They collect data from a random sample of 100 users. What type of
statistics is the researcher most likely using?
A. Descriptive statistics
B. Inferential statistics
C. Qualitative analysis
D. Experimental statistics

Answer: B. Inferential statistics (since the researcher is using a sample to make an


estimate about the entire population)

16. The probability of drawing a red card from a standard deck of cards is 0.5. What
does this probability represent?
A. The number of red cards divided by the total number of cards.
B. The ratio of red cards to the total possible outcomes in the sample space.
C. The number of successful events divided by the number of total trials.
D. The likelihood of drawing a black card.

Answer: B. The ratio of red cards to the total possible outcomes in the sample space
(since a standard deck has 52 cards, with 26 red cards, so 26/52 = 0.5)

17. A researcher wants to know the average height of students in a school with 1,000 students.
Rather than measuring all 1,000 students, the researcher decides to measure a random sample
of 50 students. Which of the following best describes this approach?
A. Convenience sampling
B. Simple random sampling
C. Stratified sampling
D. Cluster sampling

Answer: B. Simple random sampling (since each student has an equal chance of being
selected, and the sample is random)

18. A population of 500 employees is divided into three groups: managers, technicians, and
clerks. The researcher then randomly selects 50 employees from each group to participate in
the survey. What type of sampling method is being used?
A. Convenience sampling
B. Simple random sampling
C. Stratified sampling
D. Systematic sampling

Answer: C. Stratified sampling (the population is divided into subgroups, and samples are
taken from each subgroup)
19. What is the primary advantage of using random sampling over convenience sampling?
A. Random sampling guarantees a larger sample size.
B. Random sampling ensures that every individual in the population has an equal chance of
being selected.
C. Random sampling is faster and less costly.
D. Random sampling can be done without any organization.

Answer: B. Random sampling ensures that every individual in the population has an equal
chance of being selected (this reduces bias).

20. Suppose you conduct a survey by taking a random sample of 100 people to estimate the
average income of all residents in a city. If the sample has a large variation in the data, which
of the following could be true?
A. The sample is likely not representative of the population.
B. The sample size should be smaller to reduce the variation.
C. The population must have low variation in income.
D. The results will always be highly accurate regardless of the variation.

Answer: A. The sample is likely not representative of the population (large variation in the
sample may suggest bias or poor sampling techniques).

21. What is the effect of increasing the sample size on the variability (standard error) of the
sample mean?
A. Increasing the sample size decreases the variability of the sample mean.
B. Increasing the sample size increases the variability of the sample mean.
C. Increasing the sample size has no effect on the variability.
D. The effect of increasing the sample size on variability is unpredictable.

Answer: A. Increasing the sample size decreases the variability of the sample mean (the
standard error decreases as the sample size increases).

22. A survey is conducted to measure the average amount of time spent commuting to work.
The first sample includes workers from a suburban office park, while the second sample
includes workers from downtown. The results of the two samples vary significantly. What is a
likely reason for this variation?
A. The samples were not randomly selected.
B. There is inherent variability in the population.
C. The samples were too small to draw valid conclusions.
D. The two groups being sampled are likely from different subpopulations.

Answer: D. The two groups being sampled are likely from different subpopulations (workers
in suburban office parks may have different commuting times than workers in downtown
areas).

23. Which of the following sampling methods is most likely to reduce variation within the
sample?
A. Simple random sampling
B. Stratified sampling
C. Convenience sampling
D. Cluster sampling

Answer: B. Stratified sampling (by ensuring that different subgroups are proportionately
represented in the sample, stratified sampling helps reduce variation in the sample).
24. A sample of 30 students is selected from a large university to study their GPA. If the
sample has high variability in GPA, what can you conclude about the sample?
A. The sample is perfectly representative of the population.
B. The sample may not be representative of the population.
C. The sample is too large to accurately represent the population.
D. High variability indicates an error in sampling.

Answer: B. The sample may not be representative of the population (high variability could
indicate that the sample is not well selected).

25. If two different samples taken from the same population yield significantly different
results, what could explain this variation?
A. The samples were taken randomly, so differences are expected.
B. The population is highly homogeneous.
C. The sample sizes were too large.
D. The sampling methods used were likely biased or flawed.

Answer: A. The samples were taken randomly, so differences are expected (random variation
can cause differences between samples, even if the sampling method is correct).

26. A frequency table shows the following data for daily coffee consumption in a group of
people:

Which of the following insights can be drawn from the data?

A. Most people drink exactly 2 cups of coffee daily.


B. The data suggests coffee consumption is normally distributed.
C. 3 cups of coffee is the mode of the dataset.
D. Everyone drinks coffee daily.

27. In a frequency table for survey responses, you notice that one response category has a
significantly higher frequency than the others. What should you consider before interpreting
this result?
A. Whether the response category is an outlier
B. Whether the survey had biased questions or sampling
C. Whether the table was constructed correctly
D. All of the above

28. A researcher is analyzing customer satisfaction ratings on a scale of 1 to 5. Why is it


inappropriate to calculate the mean for this data?
A. The data is nominal and cannot be averaged.
B. The data is ordinal, and the mean assumes equal intervals between values.
C. The data is ratio, and calculating the mean is only valid for nominal data.
D. Satisfaction ratings always require using the median.
29. A dataset contains the following variables: "Age," "Education Level," "Income," and
"Favorite Color." Which variable is least likely to benefit from being summarized in a
frequency table, and why?
A. Age, because it is a continuous variable.
B. Education Level, because it is an ordinal variable.
C. Income, because it is a ratio variable.
D. Favorite Color, because it is a nominal variable.

30. After analyzing a frequency table, you find that one category has a cumulative relative
frequency of 95%. What does this suggest about the data?
A. This category contains most of the data values.
B. This category and the ones before it include 95% of the data values.
C. The data is skewed toward the higher end.
D. The table has an error since relative frequencies should sum to 100%.

You might also like