SAT Math Statistics: 8 Question Types
SAT Math Statistics: 8 Question Types
The IQR measures the spread of the middle 50% of a data set, providing insights into its variability and potential outliers. It is calculated by subtracting the first quartile (Q1) from the third quartile (Q3), i.e., IQR = Q3 - Q1. This reveals the range within which the central half of the data lies, disregarding the influence of extreme outliers .
To find the number x that, when added to a data set, maintains its mean, you can use the formula for the mean: Mean = (Sum of all data points + x) / (Number of data points + 1). By rearranging this equation, you can solve for x given the new desired mean and the properties of the existing data set. For example, if the numbers 3, 7, 8, 8, and x have a mean of 8, the equation would be 3 + 7 + 8 + 8 + x = 8 * 5, leading to x = 16 .
A new data point much larger than the other values will disproportionately affect the mean more than the median. This occurs because the mean is sensitive to the magnitude of all data points, hence a large value significantly increases the average. Conversely, the median, being the middle value, is less affected unless the new number changes the middle order of values. In cases where the new number is an extreme outlier, only the mean experiences significant change .
When a new number significantly larger than the existing largest value is added to a data set, the median can remain unchanged if the new number does not influence the middle order of the dataset. In an odd-sized set where it directly affects the count, it may adjust the positioning of the median, although it's less sensitive compared to the mean's response to such a change .
Adding a constant value to each data point in a data set increases the mean of the set but leaves the standard deviation unchanged. This is because standard deviation measures the dispersion of data points relative to the mean, and since each point is shifted equally, the relative dispersion does not change .
If two data sets have the same mean but different standard deviations, it means that the data set with the higher standard deviation has more variability among its data points. This means the data points are more spread out from the mean, indicating less consistency within the values. Conversely, the data set with the lower standard deviation has data points closer to the mean, indicating more consistency and less variability .
Voluntary response sampling can introduce bias because it tends to attract individuals who feel strongly about the subject, often leading to responses that are not representative of the general population. This self-selection bias can skew results toward the opinions of more vocal or passionate respondents, potentially misleading the conclusions drawn from the research .
An outlier affects the mean of a data set more than the median. This is because the mean is an arithmetic average of all data points, so an extreme value skews it in the direction of the outlier. The median, however, is the middle value when the data is ordered and is robust to extreme values, meaning it remains relatively stable when outliers are added .
Replacing several elements of a data set with much larger numbers significantly increases the mean because it is the average of all values and is affected by changes in magnitude. The median, however, might not change much unless the order of the middle values is altered significantly, as the median is solely dependent on the relative position of values .
The weighted average is calculated by multiplying the average score of each class by the number of students in that class, summing these products, and then dividing by the total number of students. For a class of 20 students with an average score of 70 and another class of 10 students with an average score of 80, the weighted average is (20*70 + 10*80) / (20 + 10) = 73.33 .