0% found this document useful (0 votes)
10 views15 pages

Understanding Population and Sample Data

A population includes all entities of interest in a study, but it is often impossible to obtain information about the entire population. Therefore, researchers take a sample, or subset, of the population to gain insights into its characteristics. A sample provides information about the larger population it was selected from.

Uploaded by

Jayson Meperaque
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
10 views15 pages

Understanding Population and Sample Data

A population includes all entities of interest in a study, but it is often impossible to obtain information about the entire population. Therefore, researchers take a sample, or subset, of the population to gain insights into its characteristics. A sample provides information about the larger population it was selected from.

Uploaded by

Jayson Meperaque
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

A population includes all of the entities of interest: people, households, machines, or whatever.

In
these situations and many others, it is virtually impossible to obtain information about all members
of the population. For example, it is far too costly to ask all potential voters which presidential
candidates they prefer. Therefore, we often try to gain insights into the characteristics of a
population by examining a sample, or subset, of the population.

A population includes all of the entities of interest in a study. A sample is a subset of the
population.
Data Sets Variables Observations
A data set is generally a A characteristic of The raw information
rectangular array of members of a collected during a study,
data where the population, such as experiment, survey, or any
columns contain height, gender, or other data-gathering
variables. salary. process.
Numerical and categorical data are two fundamental types of data that are commonly encountered
in various fields, including statistics, data science, and research. Numerical data consists of numbers
and represents measurable quantities. It can be further classified into two subtypes such as
continuous data and discrete data.

A numerical variable is discrete if it results from a count, such as the number of children. A
continuous variable is the result of an essentially continuous measurement.
It is expressed numerically, on a 1-to-5 scale. However, these numbers are really only codes for the
categories “strongly disagree,” “disagree,” “neutral,” “agree,” and “strongly agree.” There is never any
intent to perform arithmetic on these numbers; in fact, it is not really appropriate to do so. Therefore,
it is most appropriate to treat the opinion variable as categorical. Categorical data represents
categories or labels and cannot be measured in the same way as numerical data.

Nominal Ordinal
Data Data
Nominal data consists of Ordinal data has categories with a
categories with no meaningful order or ranking. However, the
inherent order or intervals between the categories are not
ranking. Examples consistent. Examples include education
include colors, gender, levels (e.g., high school, college, graduate
and types of fruit. school) or customer satisfaction ratings (e.g.,
poor, satisfactory, excellent).
The Birth of the Customer Feedback Form

The origins of customer feedback forms can be traced back to the 18th century. However, the first
modern customer feedback form was introduced in the early 20th century by a rather unexpected
industry – the automobile sector. In 1908, Henry Ford's revolutionary Model T was released, and with
it came a detachable postcard for customers to provide feedback on their driving experiences. This
early version of a customer survey was aimed at understanding customer satisfaction and improving
the driving experience for Ford's rapidly growing customer base.

Photos from Ford


There are three common measures of central tendency, all of which try to answer the basic question of which
value is most “typical.” These are the mean, the median, and the mode.

The mean is the average of all values. The most widely used measure of central tendency is the mean, or
arithmetic average. It is the sum of all the scores in a distribution divided by the number of cases. In terms of
a formula:

For Excel data sets, you can calculate the mean with the AVERAGE function.
The median is the middle observation when the data are sorted from smallest to largest.

The median can be calculated in Excel with the MEDIAN function.


The mode is the value that appears most often, and it can be calculated in Excel with the MODE function.
The mode is the value in a distribution that occurs most frequently. It is the simplest to find of the three
measures of central tendency because it is determined by inspection rather than by computation. Given the
distribution of scores.

You might also like