0% found this document useful (0 votes)
3 views32 pages

Understanding Data Measurement Types

Uploaded by

dharmaditya819
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
3 views32 pages

Understanding Data Measurement Types

Uploaded by

dharmaditya819
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

Concept of measurement scale

When you’re collecting survey data (or, really any kind of quantitative data) for your research project, you’re
going to land up with two types of data – categorical and/or numerical. These reflect different levels of
measurement.

Categorical data is data that reflect characteristics or categories (no big surprise there!). For example,
categorical data could include variables such as gender, hair colour, ethnicity, coffee preference, etc. In other
words, categorical data is essentially a way of assigning numbers to qualitative data (e.g. 1 for male, 2 for
female, and so on).

Numerical data, on the other hand, reflects data that are inherently numbers-based and quantitative in nature.
For example, age, height, weight. In other words, these are things that are naturally measured as numbers (i.e.
they’re quantitative), as opposed to categorical data (which involves assigning numbers to qualitative
characteristics or groups).

Within each of these two main categories, there are two levels of measurement:

A) Categorical data – nominal and ordinal


B) Numerical data – interval and ratio
Nominal data or scale
As we’ve discussed, nominal data is a categorical data type, so it describes qualitative
characteristics or groups, with no order or rank between categories. Examples of nominal data
include:

 Gender, ethnicity, eye colour, blood type


 Brand of refrigerator/motor vehicle/television owned
 Political candidate preference, shampoo preference, favourite meal

In all of these examples, the data options are categorical, and there’s no ranking or natural
order. In other words, they all have the same value – one is not ranked above another. So, you
can view nominal data as the most basic level of measurement, reflecting categories with no
rank or order involved.
Ordinal data / Scale

Ordinal data kicks things up a notch. It’s the same as nominal data in that it’s looking at categories,
but unlike nominal data, there is also a meaningful order or rank between the options. Here are
some examples of ordinal data:

 Income level (e.g. low income, middle income, high income)


 Level of agreement (e.g. strongly disagree, disagree, neutral, agree, strongly agree)
 Political orientation (e.g. far left, left, centre, right, far right)

As you can see in these examples, all the options are still categories, but there is an ordering or
ranking difference between the options. You can’t numerically measure the differences between
the options (because they are categories, after all), but you can order and/or logically rank them.
So, you can view ordinal as a slightly more sophisticated level of measurement than nominal.
Interval data / scale
As we discussed earlier, interval data are a numerical data type. In other words, it’s a level of
measurement that involves data that’s naturally quantitative (is usually measured in numbers).
Specifically, interval data has an order (like ordinal data), plus the spaces between measurement points
are equal (unlike ordinal data).

Sounds a bit fluffy and conceptual? Let’s take a look at some examples of interval data:
 Credit scores (300 – 850)
 GMAT scores (200 – 800)
 IQ scores
 The temperature in Fahrenheit

Importantly, in all of these examples of interval data, the data points are numerical, but the zero point is
arbitrary. For example, a temperature of zero degrees Fahrenheit doesn’t mean that there is no
temperature (or no heat at all) – it just means the temperature is 10 degrees less than 10. Similarly, you
cannot achieve a zero credit score or GMAT score.

In other words, interval data is a level of measurement that’s numerical (and you can measure the
distance between points), but that doesn’t have a meaningful zero point – the zero is arbitrary.
Long story short – interval-type data offers a more sophisticated level of measurement than nominal and
ordinal data, but it’s still not perfect.
Ratio data / scale
Ratio-type data is the most sophisticated level of measurement. Like interval data, it
is ordered/ranked and the numerical distance between points is consistent (and can be measured).
But what makes it the king of measurement is that the zero point reflects an absolute zero (unlike
interval data’s arbitrary zero point). In other words, a measurement of zero means that there is
nothing of that variable.

Here are some examples of ratio data:


 Weight, height, or length
 Length of time/duration (e.g. seconds, minutes, hours)

In all of these examples, you can see that the zero point is absolute. For example, zero seconds
quite literally means zero duration. Similarly, zero weight means weightless. It’s not some arbitrary
number. This is what makes ratio-type data the most sophisticated level of measurement.

With ratio data, not only can you meaningfully measure distances between data points (i.e. add and
subtract) – you can also meaningfully multiply and divide. For example, 20 minutes is indeed twice
as much time as 10 minutes. You couldn’t do that with credit scores (i.e. interval data), as there’s no
such thing as a zero credit score. This is why ratio data is king in the land of measurement levels.
Compare four types of scale
Concept of Scale
The concept of scale refers to the ratio of the size of something to the size of the real thing it
represents, used for measuring or representing objects, maps, or phenomena. It can be a
mathematical ratio, such as on a map or blueprint, where a scale of 1:5 means 1 unit in the drawing
represents 5 units in reality. It also refers to a standard of measurement or a system used to classify
things in science, or even a layer of complexity in social and natural systems.
In representation (maps, drawings, models)
Definition: The ratio between a drawing or model's linear dimensions and the actual dimensions of
the object.
Purpose: To accurately represent objects that are too large (like a building) or too small (like a
watch) to be shown at their actual size on paper.
Examples: A map uses scale to show the relationship between a distance on the map and the
corresponding distance on the Earth's surface.A blueprint for a house uses a scale like 1 inch to 1
foot to show the relationship between the drawing and the actual construction.
Types of representation:
Representative Fraction: Expresses scale as a ratio, e.g., 1:100.
Statement Scale: Describes the scale in words, e.g., "1 inch represents 1 mile".
Graphical Scale: Uses a bar or line marked with distance increments.
Statement Scale
A statement scale in geography is a way to express the relationship between a map's distance and the
real-world distance using a written statement, like "1 cm represents 10 km" or "1 inch equals 1 mile". It is
simple to understand but can be problematic if a map is enlarged or reduced, as the statement becomes
inaccurate.
Key characteristics
Verbal description: It uses words to describe the scale, making it easy for most people to understand.
Ratio of distances: It directly states the ratio between a distance on the map and the corresponding
distance on the ground.
Examples: 1 cm on the map represents 10 km on the ground. 1 inch on the map represents 1 mile on
the ground.
Advantages
Simplicity: It is a straightforward and intuitive method of showing scale.
Direct representation: It gives a clear idea of what a certain distance on the map means in reality.
Disadvantages
Inaccuracy with resizing: If a map is enlarged or reduced, the statement scale becomes incorrect and
requires recalculation.
Unit dependence: It can be confusing for users who are not familiar with the units of measurement
used (e.g., metric vs. imperial).
Representative Fraction (RF)
A representative fraction (RF) is a mathematical ratio that shows the relationship between a distance on a map
and the corresponding distance on the ground, using the same units for both. For example, an RF of 1:24,000
means one unit on the map represents 24,000 of the same units on the ground. This scale is versatile because it
can be applied with any unit of measurement, such as centimeters or inches.
How to understand and use an RF scale
The ratio: The RF is expressed as a fraction, usually with a colon, where the first number is the map distance
and the second is the ground distance.
Same units: Both the numerator and the denominator must be in the same unit, like inches to inches, or
centimeters to centimeters.
Example: A scale of 1:24,000 means 1 inch on the map equals 24,000 inches on the ground, or 1 cm on the map
equals 24,000 cm on the ground.
Universal application: A key advantage of the RF is its universal application, as you can work in any unit you
choose, such as metric or imperial.
Converting between RF and statement scale
From statement to RF: To convert a statement like "1 cm = 2.5 km," first convert both sides to the same unit.
Since 1 km is equal to 100,000 cm, 2.5 km is 250,00 cm. The RF would then be 1:250,000.
From RF to statement: To convert an RF like 1:1,000,000, you can convert the ground value into a more
practical unit. Since 1 km is 100,000 cm, the ground distance of 1,000,000 cm is equal to 10 km. Therefore, the
statement scale is 1 cm = 10 km.
Graphical Scale
A graphical bar scale, also called a linear scale or scale bar, is a visual representation on a map or drawing that shows
how to measure real-world distances. It looks like a small ruler, with a line divided into segments that represent ground
distances, such as kilometers or miles. This scale remains accurate even if the map is enlarged or reduced, making it
useful for measuring distances on a copied or projected map where written scales would become incorrect.

How it works
Visual representation: It's a graphic line, often with alternating colors, divided into equal sections.
Labeling: Each section is labeled with the corresponding real-world distance (e.g., 0, 1 km, 2 km).
Measurement: To find the distance between two points on a map, you measure the distance with a ruler and then
compare that measurement to the bar scale.
Advantages
Resizes with the map: Unlike a simple written scale (e.g., 1 cm = 1 km), the bar scale stays accurate even if the map is
copied or enlarged, because the entire image is scaled proportionally.
Simple and visual: It is easy to understand and interpret, making it accessible to a wide audience.
Helps in converting scales: You can easily convert a graphic scale to a verbal scale or representative fraction, and vice
versa.
Disadvantages
Single unit: A typical graphic scale is limited to a single unit of measurement. To find a distance in a different unit, you
would need to do a mathematical conversion or have a second bar scale with different units shown.
Mean
Mean

Mean = 482.5 persons/km2


Median
Median
Bar Graph
A bar graph is a graphical representation of data that uses rectangular bars of varying lengths to compare different
categories. The length or height of each bar is proportional to the value it represents, making it easy to visually
compare quantities. Bar graphs can be vertical or horizontal and are used in various fields like statistics and
business to present data clearly.
How a bar graph works
Bars: Rectangular bars are used to represent data. Their lengths are proportional to the values they represent.
Axes: A bar graph has two axes, typically a horizontal (x-axis) and a vertical (y-axis). One axis shows the categories
being compared, and the other shows the numerical values.
Comparison: The primary use is to compare discrete categories, such as the number of students who prefer
different ice cream flavors, as explained in this YouTube video.
Types of bar graphs
Vertical or Horizontal: Bars can be drawn either vertically or horizontally.
Grouped: These are used to compare multiple sets of data within the same categories, such as comparing the
budgets for two different houses, notes [Link].
Stacked: These show how a whole is divided into parts, with the bars stacked on top of each other.
Key advantages
Easy to understand: Bar graphs make it simple to see differences between categories at a glance.
Simple to create: They can be drawn by hand or with computer software.
Widely used: They are a common tool for data representation across many industries.
Expenditures per Month Under various Heads
Compound Bar Graph
A compound bar graph, also known as a grouped bar graph, compares two or more data sets at once by placing
bars side-by-side or stacking segments of bars within the same overall bar. It is useful for comparing different
categories across multiple groups, such as comparing the sales of different products in different years or showing
the population breakdown of men and women in various cities. Each bar is segmented or colored differently to
represent the various data sets for a clearer comparison.
How it works
Grouping: Bars representing different categories (e.g., different products) are grouped together for a specific data
point (e.g., a specific year).
Segmentation: Within a single bar, different segments represent the data for the various groups being compared.
For example, a single bar for the year 2023 might be split into segments for sales in the U.S., Europe, and Asia.
Comparison: Each segment is visually distinct, often using different colors or shades, making it easy to compare
values for the same sub-category across different groups.
Legibility: It's crucial to have a legend to show which color or shade corresponds to which data set (e.g., men vs.
women).
Examples
Comparing the revenue and expenditure of a company over several years.
Showing the number of students in different grades for various subjects.
Illustrating the population of men and women in different cities.
Representing the breakdown of electricity generation (e.g., thermal, hydropower) for different countries.
Gross Electricity Generation in India
Line Graph
A line graph is a type of chart that displays data as a series of points connected by straight lines, most commonly
used to show trends over a continuous interval like time. It visualizes the relationship between two variables, where
the horizontal axis (X-axis) typically represents time or an independent variable, and the vertical axis (Y-axis) shows
the corresponding values or dependent variable. These graphs make it easier to understand patterns, changes, and
relationships in data.
Key features
Data points: Individual pieces of data are represented by markers (points) on the graph.
Connecting lines: These points are connected by straight line segments to show the data's progression.
Axes: A line graph has a horizontal (X) axis and a vertical (Y) axis.
The X-axis often represents a continuous progression, such as time (days, months, years).
The Y-axis represents the numerical values or metrics for the data being tracked.
Trends: The line connecting the points makes it easy to see trends, such as increases, decreases, or fluctuations,
over time.
Common uses
Tracking changes over time: Showing how a company's sales, a person's savings, or a country's population has
changed year by year.
Comparing data sets: Using multiple lines on a single graph to compare different data sets against the same
independent variable, such as comparing the sales of two different products over the same time period.
Illustrating relationships: Showing how one variable changes in response to another.
Annual Growth of Population in India (1901-2011)
Histogram
A histogram is a graphical representation of the distribution of a dataset, using bars to show the frequency of data points that fall
within a specific range or "bin". It is useful for visualizing how often different values occur in a continuous numerical variable,
revealing patterns, trends, and outliers within the data.
How it works
Binning the data: The range of values in a dataset is first divided into a series of intervals, or "bins".
Counting frequencies: The number of data points that fall into each bin is counted to determine the frequency for that interval.
Creating the bars: A bar is drawn for each bin, where the height (or area) of the bar is proportional to the frequency of data points
within that bin.
Interpreting the chart: The resulting chart provides a quick way to see the shape of the data distribution, such as whether it's
skewed to one side or has one or more peaks.
Common uses
Statistics: To understand the distribution of a single, continuous variable like test scores, ages, or heights.
Quality control: To monitor and ensure that a manufacturing process stays within acceptable limits.
Image analysis: In digital image editors, a histogram shows the distribution of pixels by brightness, allowing users to adjust
contrast and brightness.
Finance: To analyze the momentum of a stock by showing the difference between a security's price and its moving average.
Key features
Continuous data: Unlike a bar graph which can display categorical data, a histogram is designed for continuous data.
Adjacent bars: The bars in a histogram are typically adjacent to each other to show the continuous nature of the data.
Visualizing distribution: It's an effective tool for quickly grasping the shape, central tendency, and spread of a dataset.
Histogram (equal class width) Histogram (unequal class width)

You might also like