0% found this document useful (0 votes)
24 views33 pages

Data Processing Steps for Analysis

The document discusses the process of data processing which involves preparing raw data for analysis through procedures such as editing, coding, sorting, and tabulation. It explains that the data must be checked for accuracy and organized in a way that allows for meaningful analysis through tables, graphs, and charts. The selection of data processing methods depends on factors like the type and amount of data, the desired end product, time constraints, and available resources.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PPTX, PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
24 views33 pages

Data Processing Steps for Analysis

The document discusses the process of data processing which involves preparing raw data for analysis through procedures such as editing, coding, sorting, and tabulation. It explains that the data must be checked for accuracy and organized in a way that allows for meaningful analysis through tables, graphs, and charts. The selection of data processing methods depends on factors like the type and amount of data, the desired end product, time constraints, and available resources.
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PPTX, PDF, TXT or read online on Scribd

Data Processing

= Data collected are processed in preparation for data


analysis.
= Data processing involves a variety of operations, such as,
editing, coding, encoding, sorting and tabulation.
= The selection of a particular procedure may depend on
such factors as nature and amount of the data to be
processed, the kind of end product desired, time allowed to
finish the work, and the availability of resources.
What is data processing

Data processing is a set of procedures of categorizing, organizing, and


presenting data in suitable forms that will make the data meaningful to
the researcher and suggest what statistical analysis need to be done so
that these can be correctly interpreted.

When all the data needed to answer the study objectives have
collected, the data should be processed in preparation for analysis.
Why is data processing necessary?
• In studies involving collection and analysis of quantitative
information, a great amount of information needs to be processed in
preparation for collation and analysis.

• With the availability of computers, data processing has become


simple, fast and accurate.

• Large amount of information can be stored in data files, which can be


easily retrieved and analyzed.
Through data processing:

• Data can be checked for • Tables, graphs, charts and other


completeness, consistency, and forms of data presentation can
accuracy. be easily generated.

• Coded and organized data can • Statistical analysis and


be easily and safely stored. generation of statistical outputs
can be don easily and quickly
Before starting with data processing the researcher
must decide on:

1. Whether processing will be done manually or by using a


computer;
2. What tables, graphs, charts, etc. need to be generated, and
3. What statistical manipulations will be performed.
Steps in data
processing:
There are five steps in data processing, namely;

Editing
Coding
Encoding
Creation of data files
Tabulation
Step One: Editing

• Editing is done on accomplished questionnaires/interview schedules


to discover
• omissions
• Inconsistency of responses
• Or incompleteness of information

• Errors or omissions should be remedied before data are coded.

• Editing is performed by the data collector and the field supervisor.


Tips in editing
1. Review the completed instruments immediately after interview or
administration.
Make sure all necessary questions have been asked and answered properly and legibly.
If there are omissions or inconsistencies I responses, the data collected should go back
to the respondent for clarification or additional information.

2. If the supervisor or office editor finds errors, omissions, and inconsistencies


in the instrument, he/she should clarify these with the data collector.
If the data collector cannot provide or clarify the answer, the data collector should be
advised to go back to the data source to the get the mission data or clarification of
ambiguous answers.
Illustration:

• In editing the completed questionnaire as indicated, the editor


notices two errors in the instrument.
• The first error is the inconsistency between the response for question
A.5. and the response for question A.4. since the answer in A4 is “Yes”
question A5 should have an answer. The second error is the omission
of the answer to question B.5. Every respondent should have an
answer question B5.
Research title: “Socioeconomic Characteristics of Graduate School Students in University A”
A. Respondents’ Personal characteristics
Sex 1 Male
2 Female
2. How old are you on your last 30
birthday
3. What is your civil Status 1 Single
2 Married
3 Widowed
4 Separated

4. Are you presently gainfully 0 No


working? 1 Yes
5. What is your major occupation? 0 Not working
1 Professional
2 Administration
3 Business
4 Farming/Fishing
5 Service/Communication
6 Others
7. On the average, how much do you
approximately earn from your major Php 25,000.00
occupation>

B. Educational Background
1. What undergraduate degree did 1 AB/BA
you finish? 2 Education
3 Commerce
4 Nursing
5 Engineering
6 Agriculture
7 Others

2. What graduate program are you 1 MAED


pursuing/ 2 MAN
3 MBA
4 MPA
5 MEEng
6 Others

3. Why did you enrol in graduate For further education


school?
4. Do you kike graduate school? 1 Yes
2 No
5. Why?
Step Two: Coding
• Coding is the process of converting all possible response categories to a question
to unique numerical code.

• The codes may be marked on the questionnaire/interview schedule, or written in


a specially prepared coding sheet. The codes are usually defined in a coding
manual.

• A coding sheet is a form that contains columns and rows where the code
symbols are entered.

• The column represent variables or question items, while the rows represent the
data souces, subjects, respondents, or instruments being coded.
Sampling coding sheet

R No. Sex Age Civ Work Emp Occ. Income Educ Course Like Why Why
Stat Stat grad No Yes

1 1 18 1 1 1 2 10,000 1 1 1 2 1
2 2 17 1 2 1 1 12,300 1 1 1 2 2
3 2 16 1 1 1 2 10,500 2 2 2 3 3
4 1 17 1 2 0 0 NAP 2 2 1 3 2
A Coding Manual

• Is a form which define variables and gives the codes for the categories
of responses for the questions or items in the research instrument.

• It specifies the variable number, variable name, item number in the


instrument, variable definition or description, categories of responses,
and codes for the categories.

• The coding manual is usually prepared in a tabular form.


Illustration: A coding manual for the questionnaire: Socioeconomic characteristics..
Variable No. Variable Name Quest Item No. Description Code Categories
1 Sex A.1 Sex of Respondent 1 Male
2 Female
2 Age A.2 Age of Respondent As is As is
3 Civil Status A.3 Civil Stat of Resp 1 Single
2 Married
3 Widowed
4 Work A.4 Working or not 1 Not working
2 Working
5 Empstat A.5 Employment Stat 1 Not working
2 Fulltime
3 Part-time
6 Occup A.6 Occupation of Resp 1 None/NAP
2 Professional
3 Administrator
4 Office employee
5 Farming/Fishing
6 Service provider
7 Others

7 Income A.7 Monthly income of resp As is As is


Var no. Var Name Quest Item # Description Code Categories
8 Educ B.1 Bachelors 1 AB/BA
degree/Degree
completed
9 Course B.2 Course pursued by 1 MAED
student 2 MAN
3 MB/MPA
4 MEEng
5 Others

10 Like Grd B.3 Whether or not resp like 1 Yes


graduate school 2 No
11 Whyno B.4 Reasons why 0 NAP
respondents don’t like 1 Difficult
graduate school 2 No time to study
3 Expensive
4 Others
12 Whyyes B.4 Reasons why 0 NAP
respondents like 1 Challenging
graduate school 2 Satisfying
3 Exciting
4 Others
How to prepare a coding
manual
1. Identify the variables or items in the research instrument. For example civil status
of respondent
2. Create a short name or label for each variable (8 characters or less) and give a
brief description of variable. For example, a possible short name for civil status is
“civstat”. Its description is civil status of respondent.”
3. Identify the categories of responses for each item/question and assign a unique
code for each category. (In many instruments, questions are provided with precoded
fixed alternative responses.
• For example, for variable “civstat,” the categories and their respective codes
are:
Variable Description Codes Categories
Name
civstat Civil status of 1 Single
respondents 2 Married
3 Widowed
4. For open-ended questions, get a sample of accomplished instruments and from
these, make a list of all responses to each question. Make sure that categories do
not overlap. An example of overlapping categories are: “psychological violence”
and “use of insults, sarcasm, and offensive language.”
5. Groups the answers to each item according to common characteristics and
elements, give a name to each group that captures their commonalities and assign
a unique code or symbol (number, letter or word) to each category.
Example for the open ended questions, “Why did you enrol in graduate
school?” the answers listed from a sample of completed questionnaires are listed in
one of the preceding table.
The answers have been grouped into five and each group is assigned a label
that captures the meaning of all categories in the group. A code is assigned to each
new category. Just in case there are responses that will be found in he other
questionnaires which cannot be assigned to any of the new labels, the label
“Others, specify” should be added, under which the additional categories can can
be assigned.
Responses Codes Categories
“for professional growth” 1 For professional growth
“for promotion”, and :for For promotion
salary improvement” 2

“for self-satisfaction,” and For self satisfaction


“self gratification” 3

“to learn new things” or to To learn new things and


get new ideas,” and “to be 4 ideas
updated”
“to gain friends,” “to know 5 To meet other people
more”
“Other answers that may Others, specify
be found “later” 6
Step Three: Encoding and Creating Data File
• After the raw data have been coded, data files are created.
• The coded data are entered and stored in a data sheet, then in a
computer disk, diskette or tape, to facilitate retrieval, processing, and
statistical manipulation.
• There are many softwares which can be used for this purpose, such as
the SPSS, a statistical package which is easily available.
• The coding manual, and the coding sheet or the precoded
questionnaires serve s the main references for encoding or creation of
data file.
• For manual processing, the accomplished coding sheet may already
serve as the data sheet or file.
• For computer processed data, the data file looks like a data sheet.
• The SPSS software provides instructions on how to create data files and
generate tables.
A sample coding sheet/ date file for the questionnaire
R No. Sex Age Civ Work Emp Occ. Income Educ Course Lik1e Why Why
Stat stat grad No Yes
School
1 1 18 1 1 1 2 10,000 1 1 1 2 1
2 2 17 1 2 1 1 12,300 1 1 1 2 2
3 2 16 1 1 1 2 10,500 2 2 2 3 3
4 1 17 1 2 0 0 NAP 2 2 1 3 2
5 1 18 2 2 0 0 NAP 3 3 1 3 2
6 1 17 2 1 2 2 12,000 1 1 1 4 3
7 1 18 1 1 2 1 14.050 2 4 2 2 4
8 1 19 2 2 0 0 NAP 4 3 1 3 2
9 2 17 1 1 1 2 13,000 3 5 1 3 1
10 1 18 2 1 0 0 NAP 4 1 2 3 4
Step four: Tabulation: Generating Data Summaries

• Before statistical computation and analysis are performed, initial


descriptive tables per variable must be generated.
• Tables allow the researcher to have a picture of the study population
in terms of the variables to be studied.
• The outputs of the process are single variable tables, crosstabulations,
graphs, or charts, and other outputs that will enable the research to
have a preliminary view of the findings.
• The preliminary view of the data will also allow the researcher to
identify and correct errors in coding and data entry.
Sample Tabulated data for the Questionnaire
Table 1 Distribution of graduates students by Sex

Sex Number Percent


Male 7 70
Female 3 30
Total 10 100

Table 2. Distribution of graduate students by work

Work status Number Percent


Working 6 60
Not working 4 40
Total 10 100
Table 3. Distribution of Graduate Students by Sex and Employment Status
Employment Sex
Status Total
Male Female

No. % No. % No. %


Working 4 57 2 67 6 60
Not working 3 43 1 33 4 40
Total 7 100 3 100 10 100
Table 4. Mean Age and mean Education of Respondents
Variables N Minimum Maximum Mean Std.
Deviation
Age 30 19 48 30.37 7.595
Education 30 0 8 2.43 1.547
Valid N 30
Data Analysis and Interpretation

• Data can be better appreciated and effectively used when


they have been analyzed and interpreted.

• Analysis enables the researcher to interpret the results of a


study and answer the research questions or study objectives.
What is Data Analysis
• Data analysis is a process of summarizing trends and patterns
observed in the data, determining major differentials or relationships
among variables used in the study and the application of appropriate
statistical tests on a set of data to answer the objectives of the study.

• The type of data analysis to use depends on:

• The objective of the study


• The kind of scales of measurement of the data or variables being
dealt with.
Scales of Measurement

• Understanding the measurement scales of data or variables helps


determine the type of statistics that can be used in analysing data to
answer the research objectives.

• There are four levels of measurement:


• Nominal
• Ordinal
• Interval and
• ratio
Nominal Scale
• The nominal scale has no mathematical value. It is also called a categorical
scale.

• Numbers are assigned to categories of nominal data/variables to facilitate data


processing.

• A higher number assignment does not mean a bigger value or weight.

• For example, sex is a nominal variable. Its categories, “male” and “female,” do
not have mathematical value. If number “1” is used to represent “male” and
“2” is used to represent “female”, does not mean that the “female” category
has a higher value than the “male” category. Numbers are assigned to
categories to facilitate processing.
Ordinal Scale

• An ordinal scale is a measure in which data or categories of a variability are


ordered or ranked into two or more levels of degrees, such as from low or high
or least to most.

• The distance between the first, second and the second ranks, however, is not
the same as the distance between the second and the third ranks or the
distance between third and fourth ranks.

• For example, three high school students who got the first, second and third
honors in their class obtained a general average of 94, 89 and 88, respectively.
Take note that while their ranks are consecutive, the differences in grades
between ranks are not equal.
The rank in class is an ordinal variable.
Interval Scale

• An interval scale has the characteristic of an ordinal scale, but in


addition, the distances between points in the interval scale is equal.

• For example, body temperature is considered interval scale. The


distance between a body temperature of 30 degrees Farenhiet is the
same as the distance between 40 degrees and 50 degrees.

• Body temperature does not have an absolute zero point.


Ratio Scale
• A ratio scale is almost like the interval scale, except that the ratio scale
has a real zero point.

• An example of a ratio scale is monthly income.

• Income values have equal distances between each other. For instance,
the distance between Php 1,000 and Php 3,000. similarly, the distance
between Php 5,000 and Php 10,000 is the same as the distance
between php 15,000 and Php 20,000 which is Php 5,000.
Scale Description Example
Descriptions and Examples of the Four Scales of Measurements
Nominal Categories do not have mathematical Sex: male, female
values. One is not higher or lower than the Color: Red, white, yellow
other Civil Status: Single, married
Ordinal Categories can be ranked. The difference Degree of malnutrition: 1st degree, 2nd
between the first and the second rank is degree, 3rd degree
not the same as the difference between Honor Roll: 1st, 2nd, 3rd
the second and the third ranks. Level of anger: Not angry, Angry, Very
angry

Interval The data have numerical value. The Body temperature in Farenheit: 30
distance between two points is the same, degrees, 40 degrees, 45 degrees
but there is no zero point or it may be Buisness capital (Php): 1 M, 2 M, 3 M
arbitrary.

Ratio The same as interval data but the zero No. of children: 0, 1, 2, 3, 4
point is fixed. Hrs spent in studying: 0, 5, 10
Data Analysis
• Data analysis is the process of determining the distribution of cases or
respondents under given categories of information/responses and
summarizing of trends and patterns observed in the data.

You might also like