Reasearch Methods
Reasearch Methods
ii
MODULE SUMMARY
The study of research methods provides one with the knowledge and skills needed to
solve the problems and the challenges of a fast-paced decision-making environment.
Business research course recognizes that students preparing to manage business, non-
profit and public organisations, in all areas, need training in a disciplined process for
conducting an inquiry of a management dilemma, the problem or opportunity that
requires a management decision.
The module has ten lessons covering the following areas Definition, need and
components of research Identification of research area and topic, Statement of the
problem, Literature review, Types of research approaches, Methodology design,
Sampling frame and sampling techniques, Data collection tools, design and techniques,
Data analysis methods and Report writing techniques.
During the last two decades, there has been a dramatic change in the business
environment. The trend towards complexity has increased the risk associated with
business decisions, making it more important to have a sound information base. To do
well in such an environment, one will need to understand how to identify quality
information and recognize the solid, reliable research on which one’s high-risk decisions
as a manager can be based. One also needs to know how to conduct research. Developing
these skills requires understanding the scientific method as it applies to the managerial
decision-making environment. This module therefore aims to assist students to achieve
the above objective.
Gladys J. Kimutai.
iii
TABLE OF CONTENTS
MODULE SUMMARY.....................................................................................................iii
TABLE OF CONTENTS...................................................................................................iv
LECTURE ONE: INTRODUCTION TO RESEARCH......................................................1
1.0 Lecture Objectives...................................................................................................1
1.1. Introduction..............................................................................................................1
1.2. Definition of Research.............................................................................................1
1.3. The Scientific Approach to Research......................................................................2
1.4. The Importance of Research and Decision-Making................................................2
1.5. Why Do We Do Research?......................................................................................3
1.6. Purpose of Research................................................................................................4
1.7. Why Managers need Better Information.................................................................5
1.8. Types of Research....................................................................................................5
1.9. The Value of Acquiring Research Skills.................................................................9
1.10. Definition of basic terms used in research...........................................................9
1.11. Components of Research...................................................................................10
LECTURE ONE: EXERCISES.........................................................................................11
LECTURE TWO: IDENTIFICATION OF RESEARCH AREA.....................................12
2.0 Lecture Objectives.................................................................................................12
2.1 Introduction............................................................................................................12
2.2 Identification of Research Area.............................................................................12
2.3 Identifying a research problem..............................................................................13
Activity 2.1........................................................................................................................15
2.4 Stating the Problem................................................................................................15
Activity 2.2........................................................................................................................15
2.5 Stating the Purpose................................................................................................15
Activity 2.3........................................................................................................................16
2.6 Stating the Objectives............................................................................................16
Activity 2.4........................................................................................................................16
2.7 Research Questions................................................................................................17
Activity 2.5........................................................................................................................17
2.8 Formulating Hypotheses........................................................................................17
Activity 2.6........................................................................................................................18
2.9 Assumptions and Limitations................................................................................18
Activity 2.7........................................................................................................................18
LECTURE TWO: CHAPTER EXERCISES.....................................................................18
LECTURE TWO: CHAPTER EXERCISES.....................................................................19
LECTURE THREE: LITERATURE REVIEW................................................................20
3 Lecture Objectives.....................................................................................................20
3.1 Introduction............................................................................................................20
3.2 Purpose of literature review...................................................................................20
3.3 Steps in carrying out literature review...................................................................21
3.4 Sources of literature...............................................................................................21
3.5 Tips on good review of literature...........................................................................22
Activity 3.1........................................................................................................................22
LECTURE THREE: EXERCISES...................................................................................22
iv
LECTURE FOUR: ETHICS IN RESEARCH..................................................................23
4 Lecture objectives......................................................................................................23
4.1 Introduction............................................................................................................23
4.2 Ethical treatment of participants............................................................................23
4.3 Ethics and the sponsor...........................................................................................25
4.4 Researchers and team members.............................................................................27
LECTURE FOUR EXERCISES...................................................................................28
LECTURE FIVE: RESEARCH DESIGN.........................................................................29
5 Lecture objectives......................................................................................................29
5.1 Definition of research design.................................................................................29
5.2 Essentials of Research Design...............................................................................29
5.3 Classifications of Designs......................................................................................29
5.4 Major Types of Research Design..........................................................................32
LECTURE FIVE EXERCISES.........................................................................................33
LECTURE SIX: SAMPLING...........................................................................................34
6 Lecture objectives......................................................................................................34
6.1 Introduction............................................................................................................34
Population: is a complete set of individuals, cases or objects with some observable
characteristics....................................................................................................................34
6.2 The Sample Design................................................................................................34
Factors to consider in developing a sample design...........................................................34
Steps in sampling design...............................................................................................35
Characteristics of a good sample design............................................................................35
6.3 Reasons for sampling.............................................................................................35
6.4 Characteristics of a good sample...........................................................................35
6.5 Factors that influence the sample size...................................................................36
6.6 Ways of determining the sample size:...................................................................36
6.7 Sampling procedures:............................................................................................36
6.7.1 Probability Sampling Methods...............................................................36
6.7.2 Non - Probability Sampling Methods....................................................38
LECTURE SIX EXERCISES............................................................................................39
LECTURE SEVEN: MEASUREMENT...........................................................................40
7 Lecture Objectives.....................................................................................................40
7.1 Introduction............................................................................................................40
7.2 Definition of Measurement....................................................................................40
7.3 Levels of Measurement..........................................................................................41
7.4 Sources of measurement differences.....................................................................42
7.5 Types of variables..................................................................................................43
7.6 VALIDITY AND RELIABILITY IN RESEARCH..............................................46
7.7 Reliability..............................................................................................................46
7.7.1 Ways of Assessing Reliability.......................................................................47
7.7.2 Ways of improving reliability........................................................................49
7.8 Validity..................................................................................................................49
LECTURE SEVEN EXERCISES.....................................................................................52
LECTURE EIGHT: RESEARCH INSTRUMENTS........................................................53
8 Lecture Objectives.....................................................................................................53
v
8.1 QUESTIONNAIRES.............................................................................................53
8.2 INTERVIEWS.......................................................................................................56
8.3 OBSERVATION...................................................................................................62
LECTURE EIGHT EXERCISES......................................................................................64
LECTURE NINE: DATA ANALYSIS.............................................................................65
9 Lecture Objectives.....................................................................................................65
9.1 DATA PREPARATION AND DESCRIPTION...................................................65
9.2 Descriptive Statistics.............................................................................................67
9.2.1 Measures of Central Tendency......................................................................68
9.2.2 Measures of Dispersion.................................................................................68
9.2.3 Visual Displays of Data.................................................................................69
9.3 Relational Statistics...............................................................................................69
9.4 Inferential Statistics...............................................................................................71
LECTURE NINE EXERCISES........................................................................................86
LECTURE TEN: REPORT WRITING TECHNIQUES...................................................91
10 Lecture objectives..................................................................................................91
10.1 Introduction............................................................................................................91
10.2 Writing a research proposal and research reports..................................................91
10.2.1 Prefatory items...................................................................................................91
10.2.2 Introduction........................................................................................................92
10.2.3 Literature Review..............................................................................................93
10.2.4 Methodology......................................................................................................93
10.2.5 Data analysis and Findings................................................................................94
10.2.6 Summary and conclusions.................................................................................94
10.3 Characteristics of a Good Proposal:......................................................................94
10.4 Guidelines for writing a good research report.......................................................95
LECTURE TEN EXERCISES..........................................................................................96
REFERENCES..................................................................................................................97
vi
LECTURE ONE: INTRODUCTION TO RESEARCH
1.0 Lecture Objectives
1.1. Introduction
The study of research methods provides one with the knowledge and skills needed to
solve the problems and the challenges of a fast-paced decision-making environment.
Business research course recognizes that students preparing to manage business, non-
profit and public organisations, in all areas, need training in a disciplined process for
conducting an inquiry of a management dilemma, the problem or opportunity that
requires a management decision.
During the last two decades, there has been a dramatic change in the business
environment. Emerging from a historically economic role, the business organisation has
evolved in response to the social and political mandates of national public policy,
explosive technology growth and continuing innovations in global communications.
These changes have created new knowledge needs for the manager and new publics to
consider when evaluating any decision. Other knowledge demands have arisen from
problems with mergers, trade policies, protected markets, technology transfers and
macro-economic savings –investment issues.
The trend towards complexity has increased the risk associated with business decisions,
making it more important to have a sound information base. To do well in such an
environment, one will need to understand how to identify quality information and
recognize the solid, reliable research on which one’s high-risk decisions as a manager can
be based. One also needs to know how to conduct research. Developing these skills
requires understanding the scientific method as it applies to the managerial decision-
making environment.
1
Research involves a critical analysis of existing conclusions or theories with regard to
newly discovered facts i.e. it’s a continued search for new knowledge and
understanding of the world around us.
Research is a process of arriving at effective solutions to problems through systematic
collection, analysis and interpretation of data.
Decision-Making
2
Managers make decisions every day. Ideally, such decisions would be made on the basis
of evidence thoughtfully and appropriately gathered. The more important the decisions
and their impact, the more important the research becomes. Some decisions may have
consequences resulting in considerable harm to a large number of people. Some
managers make their decisions using the following ways:
Intuition [sometimes called gut decision-making]
Randomly
Mystical or supernatural guidance
Hearsay
Authority [do as we are told or appeal to authority]
Evidence gathered by another
Evidence gathered by self or colleagues
If managers use evidence gathered by others, especially those at some distance from
them, research methods experience and knowledge is useful because it gives them rules
or guidelines helpful in evaluating the quality and utility of evidence gathered by another.
Often, managers and supervisors are eager to use invalid and unreliable evidence simply
because it is easily available. In some cases, managers want evidence that supports an
existing opinion or preference. In other cases, they want evidence with unambiguous
findings and conclusions [which is rarely found]. Managers can be impatient with the
limitations and qualifications of well-done research.
In most cases, managers want the evidence NOW and that creates a variety of problems.
Decision-makers must always face the issue of deciding now or waiting until more or
better information is available. This is called the problem of sufficiency. There may never
be enough evidence to support a difficult decision.
The manager lives in three time dimensions:
The past -- accurate sense of what was accomplished and what was not
The present -- accurate sense of what is being accomplished
The future -- what should be accomplished
Research may be used to provide evidence on the first two, which supports decisions that
will have an impact in the future.
3
at the summer reading program to see if it is cost-effective. Academic librarians may be
required to conduct research and publish as a condition of employment [publication is
certainly much more important than the research].
REPORTING is a crucial element in the conduct of research. Research depends on data. In
many organizations, a wide variety of data is routinely collected and much of that
information may be "mined" by the researcher. For example, circulation data and
acquisition data are frequently available in many libraries and information centers.
Often, research is limited by what data is relatively easily available. We need to take
advantage of data already available while encouraging managers to see the value in
collecting the right information. We need to gather evidence that answers important
questions about effectiveness and efficiency rather than just what is easily counted or has
always been counted.
"The compilation of statistical information concerning library operations is too frequently
a routine operation based on tradition and with an unclear purpose, resulting in poorly
collected data utilized at a very unsophisticated level."
We Want To Do It
Curiosity is a crucial part of the human condition. Many professionals, including
information ones, want to know more about something that interests them. Do
anthropologists use government information? How many well-know children's authors
have personal websites? Where did my family come from? Are weblog sites more likely
to be created by men or women? What do teen-aged boys read?. There is an excitement
in the discovery of new information and knowing more about some topic than anyone
else. There is joy in sharing newly gathered and previously unavailable information.
Finally, publishing or sharing information via publication or public meetings provides
visibility and recognition.
In An Ideal World
In an ideal world, research methods would be an integral part of thoughtful management
of any information agency. Better reporting would result in better data, which would
result in better decisions and a much more effective, visibly so, services to the
community.
4
1.7. Why Managers need Better Information
Global and domestic competition is more vigorous
Workers, shareholders, customers and the general public are demanding to be
included in company decision-making.
Organisations are increasingly practicing data mining and data warehousing.
Communication and measurement techniques within research have been enhanced.
The quality of theories and models to explain tactical and strategic results is
improving
The power and ease of use of today’s computers to analyze data, which help in
decision-making.
There are more variables to consider in every decision.
More knowledge exists in every field of management.
Classification by purpose
1. Basic / Pure / Fundamental Research
Basic researchers are interested in deriving scientific knowledge i.e. they are
motivated by intellectual curiosity and need to come up with a particular solution. It
focuses on generating new knowledge in order to refine or expand existing theories. It
does not consider the practical application of the findings to actual problems or
situations.
2. Applied research
5
It is conducted for the purpose of applying or testing theory and evaluating its
usefulness in solving problems. It provides data to support a theory, guide theory
revision or suggest the development of a new theory.
3. Action research
It is conducted with the primary intention of solving a specific, immediate and
concrete problem in a local setting e.g. investigating ways of overcoming water
shortage in a given area. It is not concerned with whether the results can be
generalized to any other setting.
4. Evaluation Research
It is the process of determining whether the intended results were realized.
Types of evaluation research
i. Needs assessment
A need is a discrepancy between an existing set of conditions and a desired set of
conditions. The results of needs assessment study provide the foundation for
developing new programmes and for making changes in existing ones.
ii. Formative evaluation
Helps to collect data about a programme while it is still being developed e.g. an
educational programme, a marketing strategy etc.
iii. Summative evaluation
It is done after the programme has been fully developed. It is conducted to
evaluate how worthwhile the final programme has been especially compared to
similar programmes.
6
Analyze the data
Advantages of causal-comparative study
Allows a comparison of groups without having to manipulate the independent
variables
It can be done solely to identify variables worthy of experimental investigation
They are relatively cheap.
Disadvantages of causal-comparative study
Interpretations are limited because the researcher does not know whether a
particular variable is a cause or result of a behaviour being studied.
There may be a third variable which could be affecting the established
relationship but which may not be established in the study.
3. Correlation Methods
It describes in quantitative terms the degree to which variables are related. It explores
relationships between variables and also tries to predict a subject’s score on one
variable given his or her score on another variable.
Steps in correlational research
Problem statement
Selection of subjects
Data collection
Data analysis
Advantages of the correlational method
Permits one to analyze inter-relationships among a large number of variables in a
single study.
Allows one to analyze how several variables either singly or in combination might
affect a particular phenomenon being studied.
The method provides information concerning the degree of relationship between
variables being studied.
7
Preparing the instruments
Data analysis
Purpose of survey research
i. It seeks to obtain information that describes existing phenomena by asking
individuals about their perceptions, attitudes, behaviour or values.
ii. Can be used for explaining or exploring the existing status of two or more
variables, at a given point in time.
iii. It is the most appropriate to measure characteristics of large populations.
Limitations of Survey research
i. They are dependent on the cooperation of respondents.
ii. Information unknown to the respondents cannot be tapped in a survey e.g.
amount saved per year
iii. Requesting information which is considered secret and personal, encourages
incorrect answers.
iv. Surveys cannot be aimed at obtaining forecasts of things to come.
Historical research
Involves the study of a problem that requires collecting information from the past.
Purpose of Historical Research
Aims at arriving at conclusions concerning causes, effects or trends of past
occurrences that may help explain present events and anticipate future events.
Attempts to interpret ideas or events that had previously seemed unrelated.
Synthesizes old data or merges old data with new historical facts that the
researcher or other researchers have discovered.
To reinterpret past events that have been studied.
Observational Research
The current status of a phenomenon is determined not by asking but by observing.
This helps to collect objective information.
Steps
Selection and definition of the problem.
Sample selection.
Definition of the observational information.
Recording observational information
Data analysis and interpretation.
8
1. Non-participant observation
The observer is not directly involved in the situation to be observed.
2. Naturalistic Observation
Behaviour is studied and recorded as it normally occurs.
3. Simulation observation.
The researcher creates the situation to be observed and tells subjects to be observed
what activities they are to engage in. Disadvantage – the setting is not natural and the
behaviour exhibited by the subjects may not be the behaviour that would occur in a
natural setting.
4. Participant observation
The observer becomes part of or a participant in the situation. May not be ethical
5. Case studies
A case study is an in-depth investigation of an individual, group, institution or
phenomenon. It aims to determine factors and relationships among the factors that
have resulted in the behaviour under study.
6. Content analysis
It involves observation and detailed description of objects, items or things that
comprise the sample. The purpose is to study existing documents such as books,
magazines in order to determine factors that explain a specific phenomenon.
Steps
Decide on the unit of analysis
Sample the content to be analyzed
Coding
Data analysis
Compiling results and interpretations.
Advantages
Researchers are able to economize in terms of time and money.
Errors that arise during the study are easier to detect and correct.
The method has no effect on what is being studied.
Disadvantages
It is limited to recorded communication.
It is difficult to ascertain the validity of the data.
1.9. The Value of Acquiring Research Skills
Managers in today’s organisations require research skill in order to:-
To gather more information before selecting a course of action
To do a high-level research study
To understand research design
To evaluate and resolve a current management dilemma
To establish a career as a research specialist
9
Sampling: It is the process of selecting a number of individuals for a study in such a
way that the individuals selected represent the population.
Variable: It is a measurable characteristic that assumes different values among the
subjects. They can be dependent, independent, intervening, confounding or
antecedent variables.
Data: refers to all information a researcher gathers for his or her study. Can be
secondary data or primary data.
Parameter: It is a characteristic that is measurable and can assume different values in
the population.
Statistics: it is the science of organizing, describing and analyzing data. Descriptive
and inferential statistics.
Objective: it refers to the specific aspects of the phenomenon under study that the
researcher desires to bring out at the end of the research study.
Literature review: It involves locating, reading and evaluating reports of previous
studies, observations and opinions related to the planned study.
Hypothesis: It is a researcher’s anticipated explanation or opinion regarding the
result of the study.
Theory: It is a set of concepts or constructs and the interrelations that are assumed to
exist among those concepts. It provides the basis for establishing the hypothesis to be
tested in the study.
A construct is an image or idea specifically invented for a given research and/or
theory-building purpose
A concept is a bundle of meanings or characteristics associated with certain events,
objects, conditions, situations, and behaviors. Concepts have been developed over
time through shared usage
10
LECTURE ONE: EXERCISES
1.1 Define the term research.
1.2 Define the term business research.
1.3 Explain the need for a having a scientific approach to research.
1.4 Differentiate between deductive and inductive explanation.
1.5 Explain the importance of research to a business organisation.
1.6 Describe the various components of research.
11
LECTURE TWO: IDENTIFICATION OF RESEARCH AREA
2.0 Lecture Objectives
2.1 Introduction
Research originates in the decision process. A manager needs specific information for
setting objectives, defining tasks, finding the best strategy by which to carry out the tasks
or judging how well the strategy is being implemented.
The manager must proceed from the management dilemma to the management question
to proceed with the research process. The management question restates the dilemma in
question form e.g.
What should be done to reduce employee turnover?
What should be done to reduce costs?
Management questions are numerous but they can be categorized to:
Choice of purpose or objectives
Generation and evaluation of solutions
Troubleshooting or control situation
Choice of purpose or objectives:
The general question is “What do we want to achieve?”
12
Example
What goals should XYZ Co try to achieve in its next round of labour
negotiations?
Generation and evaluation of solutions:
The general question is “How can we achieve the ends we seek?”
Examples:
How can we achieve our five-year goal of doubled sales and net profits?
What should be done to reduce post purchase service complaints?
The research process starts by formulating a research problem that can be investigated
through research procedures.
13
Subject which is overdone should not be normally chosen, for it will be a difficult
task to throw any new light in such a case.
Controversial subject should not become the choice of an average researcher.
Too narrow or too vague problems should be avoided.
The subject selected for research should be familiar and feasible so that the related
research material or sources of research are within one’s reach.
The importance of the subject, the qualifications and the training of a researcher, the
costs involved and the time factor must be considered.
The selection of a study must be preceded by a preliminary study.
Defining the problem involves the task of laying down boundaries within which a
researcher shall study the problem with a predetermined objective in view. The following
steps can be followed:-
Statement of the problem in a general way
Understanding the nature of the problem: Understand the origin and nature of the
problem e.g. by discussing it with those who raised it in order to find out how the
problem originally came about. The researcher should keep in view the environment
within which the problem is to be studied and understood.
Surveying the available literature: the researcher must be well conversant with
relevant theories in the field, reports and records as also all other relevant literature.
Developing ideas through discussions:
Rephrasing the research problem: Its putting the research problem in as specific
terms as possible so that it may become operationally viable and may help in the
development of working hypotheses.
The following should also be observed when defining a research problem:
Technical terms and words or phrases with special meanings used in the statement
of the problem, should be clearly defined.
Basic assumptions or postulates if any relating to the research problem should be
clearly stated.
A straight forward statement of the value of the investigation should be provided.
The suitability of the time-period and the sources of data available must also be
considered by the researcher in defining the problem.
The scope of the investigation or the limits within which the problem is to be
studied must be mentioned explicitly in defining a research problem.
14
(f) The media
(g) Personal experiences.
Example
Assuming that we want to carry out a study, we have to start by identifying the broad
area. Our broad area in our case can be Finance or management science. . We can
identify a specific problem that will form the basis of the study like inventory
management. Since we have many inventory control techniques, we can either analyse all
the techniques or just narrow down to one technique like the just in time. The research
topic can be formulated as “An analysis of the factors affecting the application of Just-In-
Time inventory control technique: A Case of the agro-based manufacturing firms within
Nairobi”.
Activity 2.1
Identify a researcheable problem from your area of specialization
(Marketing, Human resource, Finance, Accounting or
Management Science) and formulate a research topic.
Have a write up of the background of that problem.
15
Purpose
The purpose of this study is to investigate the various inventory control techniques used
in manufacturing firms and to analyze the extent of the use of JIT.
In stating the purpose of the study, the researcher should choose the right words to
convey the focus of the study effectively. Use of subjective or biased words or sentences
should be avoided.
Examples
Biased Neutral
To show To determine
To prove To compare
To confirm To investigate
To verify To differentiate
To check To explore
Activity 2.3
Example
Objectives
To find out the inventory control techniques used in manufacturing firms in
Kenya.
To assess the extent of the application of JIT inventory control techniques in
manufacturing firms in Kenya.
To determine the factors characterizing manufacturing operations that potentially
inhibit the use of JIT.
To suggest ways of improving the application of JIT inventory control techniques
in manufacturing firms in Kenya.
Activity 2.4
16
2.7 Research Questions
The way in which one structures the research questions sets the direction for the project.
A management problem or opportunity can be formulated as a hierarchical sequence of
questions. At the most general level is the management dilemma. This is translated into a
management question and then into a research question – the major objective of the
study. The research questions can further be expanded into investigative questions.
Example:
Which inventory control techniques do manufacturing firms use?
Which factors inhibit the application of JIT within the manufacturing firms?
What are the requirements needed before the implementation of JIT?
How can the application of JIT inventory control techniques be improved in
manufacturing firms?
Activity 2.5
Purpose of hypothesis
According to Mugenda and Mugenda (2003), the purpose of hypothesis in research are:
1. It provides direction by bridging the gap between the problem and the evidence
needed for its solution.
2. It ensures collection of the evidence necessary to answer the question posed in the
statement of the problem.
3. It enables the investigator to assess the information he or she has collected from the
standpoint of both relevance and organisation.
4. It sensitizes the investigator to certain aspects of the situation that are relevant
regarding the problem at hand.
5. It permits the researcher to understand the problem with greater clarity and use the
data to find solutions to problems.
6. It guides the collection of data and provides the structure for their meaningful
interpretation in relation to the problem under investigation.
7. It forms the framework for the ultimate conclusions as solutions.
17
They should be empirical statements; never normative or value statements about
what should or should not be.
A hypothesis should describe a general phenomena not a particular occurrence.
A good hypothesis should be plausible. There should be some logical reason for
thinking it possible.
A good hypothesis is specific. The concepts used are clearly defined. An example
of a bad hypothesis is to say that there is a relationship between personality and
political attitudes. Which personality type? What attitudes? A good hypothesis is
more specific, e.g., People who feel alienated are not likely to have a strong trust
in government.
A good hypothesis is testable. There must be evidence that is obtainable which
will indicate whether the hypothesis is correct or not.
Examples
The type of product produced and sold determines the inventory control technique
used by a firm.
Instability of demand and Supplier unreliability inhibits the effective application of
Just in time technique.
Types of hypotheses
Null hypothesis (Ho): The null hypothesis is a statement about the value of a
population parameter. It should be stated as “ There is no significant difference
between ……………”. It should always contain an equal sign.
Alternate Hypothesis (HA): The alternate hypothesis is a statement that is accepted
if sample data provide enough evidence that the null hypothesis is false.
Activity 2.6
Activity 2.7
Enumerate the assumptions and limitations for your study if any.
18
LECTURE TWO: CHAPTER EXERCISES
2.1 Describe some of the factors that affect the scope of a study.
2.2 Differentiate between the null and the alternative hypothesis.
2.3 Explain the characteristics of a good problem statement.
2.4 Explain the characteristics of a good objective.
2.5 Explain the purpose of hypothesis in research
2.6 Discuss the characteristics of a good hypothesis.
2.7 Differentiate between an assumption and a limitation in research.
19
LECTURE THREE: LITERATURE REVIEW
3 Lecture Objectives
3.1 Introduction
The review of literature involves the systematic identification, location and analysis of
documents containing information related to the research problem being investigated. It
should be extensive and thorough because it is aimed at obtaining detailed knowledge of
the topic being studied. Knowledge is cumulative: every piece of research will contribute
another piece to it. That is why it is important to commence all research with a review of
the related literature or research, and to determine whether any data sources exist already
that can be brought to bear on the problem at hand.
The literature review should provide the reader with an explanation of the theoretical
rationale of the problem being studied as well as what research has already been done and
how the findings relate to the problem at hand. The quality of the literature being
reviewed must be carefully assessed. Not all published information is the result of good
research design, or can be substantiated. Indeed, a critical assessment as to the
appropriateness of the methodology employed can be part of the literature review.
According to Mugenda and Mugenda (2003), the purpose of literature review is to:
1. To determine what has already been done related to the research problem being
studied. This will help the researcher to:
Avoid unnecessary and unintentional duplication.
Form the framework within which the research findings are to be interpreted.
Demonstrate his or her familiarity with the existing body of knowledge.
2. Helps reveal the strategies, procedures and measuring instruments that have been
found useful in investigating the problem in question. This will help the researcher to:
Avoid mistakes that have been made by other researchers
Benefit from other researcher’s experiences
Clarify how to use certain procedures, which one may only have learned in
theory.
20
3. Helps to suggest other procedures and approaches, which will help, improve the
research study.
4. Familiarizes the researcher with previous studies, which facilitates interpretation of
the results of the study. If there is a contradiction, the literature review might provide
rationale for the discrepancy.
5. It helps the researcher to limit the research problem and to define it better.
6. Helps to determine new approaches and stimulates new ideas. The researcher may be
alerted to research possibilities, which have been overlooked in the past.
7. Approaches that have been proved to be futile will be revealed through literature
review.
8. Specific suggestions and recommendations for further research can be found by
reviewing literature.
9. It pulls together, integrates and summarizes what is known in an area. Thus helping to
reveal gaps in information and areas where major questions still remain.
Examples
Scholarly journals Periodicals
Theses and dissertations The Africana section of the library
Government documents Reference section of the library
Papers presented at conferences Grey literature
Books Inter-library loan
References quoted in books The British lending library
International indices The internet
Abstracts Microfilm
21
3.5 Tips on good review of literature
Do not conduct a hurried review for fear of overlooking important studies.
Do not rely too heavily on secondary sources.
Check daily newspapers as they contain very educative, current information.
Copy the references correctly in the first place so as to avoid the frustration of trying
to retrace a reference later.
Do not only concentrate on findings, check on methodology and measurement of
variables.
Activity 3.1
22
LECTURE FOUR: ETHICS IN RESEARCH
4 Lecture objectives
4.1
Int
By the end of the lecture, the students should be able to:-
Define the term ethics.
Discuss the ethical issues concerning the researcher and
(a) The respondents
(b) The sponsor
(c) The team players
roduction
Ethics are norms or standards of behaviour that guide moral choices about our behaviour
and our relationship with others. Ethics differ from legal constraints, in which generally
accepted standards have defined penalties that are universally enforced. The goal of
ethics in research is to ensure that no one is harmed or suffers adverse consequences from
research activities.
(a) Benefits
Whenever direct contact is made with a respondent, the researcher should discuss the
study benefits, being careful to neither overstate nor understate the benefits. An
interviewer should begin an introduction with his or her name, the name of the research
organisation and a brief description of the purpose and benefits of the research. This puts
the respondent at ease, lets them know to whom they are speaking and motivates them to
answer questions truthfully. Inducements to participate, financial or otherwise, should
not be disproportionate to the task or presented in a fashion that results in coercion.
Deception occurs when the respondents are told only part of the truth or when the truth is
fully compromised. The benefits to be gained by deception should be balanced against
23
the risks to the respondents. When possible, an experiment or interview should be
designed to reduce reliance on deception. In addition, the respondent’s rights and well-
being must be adequately protected. In instances where deception in an experiment could
produce anxiety, a subject’s medical condition should be checked to ensure that no
adverse physical harm follows.
24
Restricting access to data instruments where the respondent is identified.
Nondisclosure of data subsets.
Researchers should restrict access to information that reveals names, telephone numbers,
address or other identifying features. Only researchers who have signed nondisclosure,
confidentiality forms should be allowed access to the data. Links between the data or
database and the identifying information file should be weakened. Individual interview
response sheets should be inaccessible to everyone except the editors and data entry
personnel.
Occasionally, data collection instruments should be destroyed once the data are in a data
file. Data files that make it easy to reconstruct the profiles or identification of individual
respondents should be carefully controlled. For very small groups, data should not be
made available because it is often easy to pinpoint a person within the group. Employee-
satisfaction survey feedback in small units can be easily used to identify an individual
through descriptive statistics.
Privacy is more than confidentiality. A right to privacy means one has the right to refuse
to be interviewed or to refuse to answer any question in an interview. Potential
participants have a right to privacy in their own homes, including not admitting
researchers and not answering telephones. They have the right to engage in private
behaviour in private places without fear of observation. To address these rights, ethical
researchers can do the following:-
Inform respondents of their right to refuse to answer any questions or participate in
the study.
Obtain permission to interview respondents
Schedule field and phone interviews.
Limit the time required for participation.
Restrict observation to public behaviour only.
(a) Confidentiality
Sponsors have a right to several types of confidentiality including sponsor nondisclosure,
purpose nondisclosure and findings nondisclosure.
Sponsor nondisclosure: Companies have a right to dissociate themselves from the
sponsorship of a research project. Due to the sensitive nature of the management
dilemma or the research question, sponsors may hire an outside consulting or
research firm to complete research projects. this is often done when a company is
testing a new product idea, to avoid potential consumers from being influenced by
the company’s current image or industry standing. If a company is contemplating
entering a new market, it may not wish to reveal its plans to competitors. In such
25
cases, it is the responsibility of the researcher to respect this desire and device a
plan to safeguard the identity of the sponsor.
Purpose nondisclosure: It involves protecting the purpose of the study or its
details. A research sponsor may be testing a new idea that is not yet patented and
may not want the competitor to know his plans. It may be investigating employee
complaints and may not want to spark union activity. The sponsor might also be
contemplating a new public stock offering, where advance disclosure would spark
the interest of authorities or cost the firm thousands of shillings.
Findings nondisclosure: If a sponsor feels no need to hide its identity or the
study’s purpose, most sponsors want research data and findings to be confidential,
at least until the management decision is made.
The ethical course often requires confronting the sponsor’s demand and taking the
following actions: -
Educating the sponsor on the purpose of research
Explain the researcher’s role in fact finding versus the sponsor’s role in decision-
making.
26
Explain how distorting the truth or breaking faith with respondents leads to future
problems
Failing moral suasion, terminate the relationship with the sponsor.
27
LECTURE FOUR EXERCISES
4.1 Define the term ethics.
4.2 Discuss the ethical issues concerning the researcher and
(a) The respondents
(b) The sponsor
(c) The team players
28
LECTURE FIVE: RESEARCH DESIGN
5 Lecture objectives
29
Causal
The time dimension Cross-sectional
Longitudinal
The topical scope – breath and depth Case
of the study Statistical study
The research environment Field setting
Laboratory research
Simulation
The participants perceptions of Actual routine
research activity Modified routine
30
4. Purpose of the study
Descriptive study: it is a research that is concerned with finding out who, what,
where, when, or how much.
Causal study: It is concerned with learning why i.e. how one variable produces
changes in another. It tries to explain the relationships among variables.
Case studies: they place more emphasis on a full contextual analysis of fewer
events or conditions and their interrelations. Although hypotheses are often used,
the reliance on qualitative data makes support or rejection more difficult. An
emphasis on detail provides valuable insight for problem solving, evaluation and
strategy. This detail is secured from multiple sources of information. It allows
evidence to be verified and avoids missing data.
8. Participants’ perceptions
The usefulness of a design may be reduced when people in a disguised study perceive
that research is being conducted. Participants’ perceptions influence the outcomes of the
research in subtle ways. There are three levels of perception:
Participants perceive no deviations from everyday routines
Participants perceive deviations, but as unrelated to the researcher.
Participants perceive deviations as researcher-induced.
In all research environments and control situations, researchers need to be vigilant to
effects that may alter their conclusions. Participant’s perceptions serve as a reminder to
31
classify one’s study by type, to examine validation strengths and weaknesses and to be
prepared to qualify results accordingly.
Despite its obvious value, researchers and managers give exploration less attention that it
deserves. Exploration is sometimes linked to old biases about qualitative research i.e.
subjective ness, non-representativeness and non-systematic design.
When we consider the scope of qualitative research, several approaches are adaptable for
exploratory investigations of management questions:
In-depth interviewing – usually conversational rather than structured.
Participant observation – to perceive first hand what participants in the setting
experience
Films, photographs and videotapes – to capture the life of the group under study.
Case studies – for an in-depth contextual analysis of a few events or conditions
Document analysis – to evaluate historical or contemporary confidential or public
records, reports, government documents and opinions.
Where these approaches are combined, four exploratory techniques emerge with wide
applicability for the management researcher: -
i. Secondary data analysis
ii. Experience surveys
iii. Focus groups
iv. Two-stage designs
An exploratory research is finished when the researchers have achieved the following:
Established the major dimensions of the research task
Defined a set of subsidiary investigative questions that can be used as a guide to a
detailed research design.
Developed several hypotheses about possible causes of a management dilemma.
Learned that certain other hypotheses are such remote possibilities that they can be
safely ignored in any subsequent study.
Concluded additional research is not needed or is not feasible.
32
It is the process of collecting data in order to test hypotheses or to answer questions
concerning the current status of the subjects in the study. It determines and reports the
way things are. It attempts to describe such things as possible behaviour, attitudes, values
and characteristics.
(c) Causal Research
It is used to explore relationships between variables. It determines reasons or causes
for the current status of the phenomenon under study. The variables of interest cannot
be manipulated unlike in experimental research.
Advantages of causal study
Allows a comparison of groups without having to manipulate the independent
variables
It can be done solely to identify variables worthy of experimental investigation
They are relatively cheap.
Disadvantages of causal study
Interpretations are limited because the researcher does not know whether a
particular variable is a cause or result of a behaviour being studied.
There may be a third variable which could be affecting the established
relationship but which may not be established in the study.
33
LECTURE SIX: SAMPLING
6 Lecture objectives
6.1 Introduction
The method section of a research study describes the procedures that are to be followed
in conducting the study. The techniques of obtaining data are developed.
It refers to the techniques of the procedure the researcher would adopt in selecting items
for the sample.
Factors to consider in developing a sample design
1. Type of universe; finite or infinite
2. Sampling unit; geographic: state, district or village, construction unit: house, flat.
Social unit: family, club, school or individual.
3. Source list: sampling frame- contains all the names of all items of a universe. The list
should be comprehensive, correct, reliable and appropriate.
4. The size of the sample. Should be efficient, representative, reliable and flexible.
5. Parameters of interest
6. Budgetary constraint
7. Sampling procedure.
Criteria for selecting a sampling procedure
Two costs are involved in a sampling analysis i.e. the cost of collecting the data and the
cost of an incorrect inference resulting from the data. Two causes of incorrect inferences
are systematic bias and sampling error. A systematic bias results from errors in the
sampling procedures and it cannot be reduced or eliminated by increasing the sample
size. Systematic bias is the result of the following factors:-
34
Inappropriate sampling frame
Defective measuring device
Non-respondents
Indeterminancy principle – individuals act differently when kept under observation.
Natural bias in reporting data e.g. government tax – downward bias, social
organizations – upward bias.
Sampling errors are the random variations in the sample estimates around a true
population parameter. It decreases with the increase in the size of the sample and it
happens to be of a smaller magnitude in case of a homogenous population. While
selecting a sampling procedure, the researcher must ensure that the procedure causes a
relatively small sampling error and helps to control the systematic bias in a better way.
35
6.5 Factors that influence the sample size
Dispersion / variance: The greater the dispersion or variance within the population,
the larger the sample must be to provide estimation precision.
Precision of the estimate: the greater the desired precision of the estimate, the larger
the sample must be.
Interval range: The narrower the interval range, the larger the sample must be.
Confidence level: The higher the confidence level in the estimate, the larger the
sample must be.
Number of subgroups: The greater the number of subgroups of interest within a
sample, the greater the sample size must be, as each subgroup must meet minimum
sample size requirements.
If the calculated sample size exceeds 5% of the population, sample size may be
reduced without sacrificing precision.
where
n = the desired sample size
z = the standard normal deviate at the required confidence level.
P = the proportion in the target population estimated to have characteristics being
measured.
Q= 1–p
D = the level of statistical significance set.
If the target population is less than 10,000 the following formula is used to determine the
sample size;
Where
nf = the desired sample size( when the population is less than 10,000)
n = the desired sample size( when the population is greater than 10,000)
N = the estimate of the population size.
36
Types of Probability sampling methods
1. Simple Random Sampling:
A sample is selected so that each item or person in the population has the same
chance of being included.
Advantages
Easy to implement with automatic dialing and with computerized voice
response systems.
Disadvantages
Requires a listing of population elements.
Takes more time to implement
Uses larger sample sizes
Produces larger errors
Expensive
2. Systematic Random Sampling:
The items or individuals of the population are arranged in some manner. A random
starting point is selected and then every kth member of the population is selected for
the sample.
Advantages
Simple to design
Easier to use than the simple random.
Easy to determine sampling distribution of mean or proportion.
Less expensive than simple random.
Disadvantages
Periodicity within the population may skew the sample and results.
If the population list has a monotonic trend, a biased estimate will result based
on the start point.
Disadvantages
Increased error will result if subgroups are selected at different rates
Expensive especially if strata on the population have to be created.
4. Cluster Sampling:
37
The population is divided into internally heterogeneous subgroups and some are
randomly selected for further study. It is used when it is not possible to obtain a
sampling frame because the population is either very large or scattered over a large
geographical area. A multi-stage cluster sampling method can also be used.
Advantages
Provides an unbiased estimate of population parameters if properly done.
Economically more efficient than simple random.
Lowest cost per sample, especially with geographic clusters.
Easy to do without a population list.
Disadvantages
More error (Lower statistical efficiency) due to subgroups being homogeneous
rather the heterogeneous.
Advantage
Widely used by pollsters, marketers and other researchers.
Disadvantages
It gives no assurance that the sample is representative of the variables being studied.
The data used to provide controls may be outdated or inaccurate.
There is a practical limit on the number of simultaneous controls that can be applied
to ensure precision.
Since the choice of subjects is left to field workers, they may choose only friendly
looking people.
38
identified who in turn identify others. Commonly used in drug cultures, teenage gang
activities, Mungiki sect, insider trading, Mau Mau etc.
Sampling error
It’s the difference between a sample statistic and its corresponding population
parameter. The sampling distribution of the sample means is a probability distribution
of possible sample means of a given sample size.
39
LECTURE SEVEN: MEASUREMENT
7 Lecture Objectives
7.1 Introduction
While people measure things casually in daily life, research measurement is more precise
and controlled. In measurement, one settles for measuring properties of the objects rather
than the objects themselves. An event is measured in terms of its duration i.e. what
happened during it, who was involved, where it occurred etc. Measurement is the basis
for all systematic inquiry because it provides us with the tools for recording differences in
the outcome of variable change.
40
often guided by the theoretical framework, perspective, or approach the researcher is
committed to. For example, a researcher operating from within a Marxist framework
would have quite different conceptual definitions for a hypothesis about social class and
crime than a non-Marxist researcher. That's because there are strong value positions in
different theoretical perspectives about how some things should be measured.
Anything that can be measured falls into one of the four types;
The higher the level of measurement, the more precision in measurement; and
Nominal
Ordinal
Interval
41
Ratio
The nominal level of measurement describes variables that are categorical in nature.
The characteristics of the data you're collecting fall into distinct categories. If there are a
limited number of distinct categories (usually only two), then you're dealing with a
dichotomous variable. If there are an unlimited or infinite number of distinct
categories, then you're dealing with a continuous variable. Nominal variables include
demographic characteristics like sex, race, and religion.
The ordinal level of measurement describes variables that can be ordered or ranked in
some order of importance. It describes most judgments about things, such as big or little,
strong or weak. Most opinion and attitude scales or indexes in the social sciences are
ordinal in nature.
The interval level of measurement describes variables that have more or less equal
intervals, or meaningful distances between their ranks. For example, temperature, time,
The ratio level of measurement describes variables that have equal intervals and a fixed
zero (or reference) point. It is possible to have zero income, zero education, and no
involvement in crime, but rarely do we see ratio level variables in social science since it's
almost impossible to have zero attitudes on things, although "not at all", "often", and
"twice as often" might qualify as ratio level measurement.
Advanced statistics require at least interval level measurement, so the researcher always
strives for this level, accepting ordinal level (which is the most common) only when they
have to. Variables should be conceptually and operationally defined with levels of
measurement in mind since it's going to affect how well you can analyze your data later
on.
42
(d) The data collection instrument: a defective instrument can cause distortion in two
major ways:
It can be too confusing and ambiguous e.g. the use of complex words,
leading questions, ambiguous meanings, multiple questions.
Leads to poor selection from the universe of content items. Seldom does
the instrument explore all the potentially important issues.
Since absolute control of extraneous variables is not possible in any study, results are
interpreted on the basis of degrees of confidence rather than certainty.
43
Once the major extraneous variables are identified, the researcher can control them by:-
i. Building the extraneous variable into the study: i.e. including it as an independent
variable. E.g. in determining the effect of alcohol on reaction time, sex may
influence reaction time. Therefore, sex can be introduced as an independent
variable. Using regression, one can measure the effect of alcohol on reaction time,
controlling sex.
ii. Include them in the study but only at one level e.g. time is the dependent variable,
alcohol level - the independent and sex the extraneous variable. Sex can be
controlled by sampling only females or males of a given age. The disadvantage of
this method is that generalizations are limited to a smaller population.
iii. By removing the effects of the extraneous variables by statistical procedures i.e.
by siphoning its effects on the dependent variable. This can be done by:
Analysis of co-variance
Partial correlation.
Extraneous variables
They are those variables that affect the outcome of a research study either because the
researcher is not aware of their existence or if the researcher is aware, she or he has no
control over them.
Intervening variables
They are a special case of extraneous variables. The difference between the intervening
and extraneous variables is in the assumed relationship among the variables. With an
extraneous variable, there is no causal link between the independent and dependent
variable, but they are independently associated with a third variable – the extraneous
variable. An intervening variable is recognized as being caused by the independent
variable and as being a determinant of the dependent variable.
44
Independent intervening dependent
The choice of the right intervening variables helps one not only to determine accurately
the total effects of an independent variable on the dependent variable but also partition
the total effects into direct and indirect.
Antecedent variables
They do not interfere with the established relationship between an independent and
dependent variable but clarifies the influence that precedes such a relationship.
The variables including the antecedent variable must be related in some logical
sequence.
When the antecedent variable is controlled for, the relationship between the
independent and the dependent variables should not disappear. Rather it should be
enhanced.
When the independent variable is controlled for or its influence removed, there
should not be any relationship between the antecedent variable and the dependent
variable.
e.g. political stability – attracts investors – increased job opportunities – high standards of
living – reduction of poverty.
Suppressor variables
Sex and age – males 25-30 yrs – no relationship. But when the subjects are asked to press
a button when a red flash of light appears, there is a difference. Therefore the suppressor
variable is the colour of the flash light.
45
Distorter variables
It is a variable that converts what was thought of as a positive relationship into a negative
relationship and vice-versa. Its effects lead a researcher into drawing erroneous
conclusions from the data. When the distorter variable is controlled, a true relationship is
obtained. Consideration of distorter variables in a study reduces the chances of making a
type I (rejecting a true null hypothesis) or type two error (accepting a false null
hypothesis).
They are commonly used in testing hypothesized causal models. Path analysis ( a
procedure that tests causal links among several variables) is often used in testing the
validity of causal relationships in a theory or model.
A C
B D
A and B are called exogenous variables. They lack hypothesized causes in the model.
7.7 Reliability
Reliability is the extent to which an experiment, test, or any measuring procedure yields
the same result on repeated trials. Without the agreement of independent observers able
to replicate research procedures, or the ability to use research tools and procedures that
yield consistent measurements, researchers would be unable to satisfactorily draw
conclusions, formulate theories, or make claims about the generalizability of their
research. In addition to its important role in research, reliability is critical for many parts
of our lives, including manufacturing, medicine and sports. Reliability is such an
important concept that it has been defined in terms of its application to a wide range of
activities.
46
Reliability is influenced by random error. Random error is the deviation from a true
measurement due to factors that have not effectively been addressed by the researcher. As
random error increases, reliability decreases.
Test-Retest
Equivalent form
Internal consistency
Interrater reliability
1. The Test-Retest technique
It involves administering the same instruments twice to the same group of subjects, but
after some time. Stability reliability (sometimes called test, re-test reliability) is the
agreement of measuring instruments over time. To determine stability, a measure or test
is repeated on the same subjects at a future date. Results are compared and correlated
with the initial test to give a measure of stability.
2. Equivalent form
Equivalent reliability is the extent to which two items measure identical concepts at an
identical level of difficulty. Equivalency reliability is determined by relating two sets of
test scores to one another to highlight the degree of relationship or association. In
47
quantitative studies and particularly in experimental studies, a correlation coefficient,
statistically referred to as r, is used to show the strength of the correlation between a
dependent variable (the subject under study), and one or more independent variable,
which are manipulated to determine effects on the dependent variable. An important
consideration is that equivalency reliability is concerned with correlational, not causal,
relationships.
Two instruments are used. Specific items in each form are different but they are designed
to measure the same concept. They are the same in number, structure and level of
difficulty e.g. TOEFL, GRE
Advantages
Estimates the stability of the data as well as the equivalence of the items in the two
forms
Disadvantages
Difficulty in constructing two tests, which measure the same concept (time and
resources).
4. Interrater reliability
Interrater reliability is the extent to which two or more individuals (coders or raters)
agree. Interrater reliability addresses the consistency of the implementation of a rating
system.
A test of interrater reliability would be the following scenario: Two or more researchers
are observing a high school classroom. The class is discussing a movie that they have just
48
viewed as a group. The researchers have a sliding rating scale (1 being most positive, 5
being most negative) with which they are rating the student's oral responses. Interrater
reliability assesses the consistency of how the rating system is implemented. For
example, if one researcher gives a "1" to a student response, while another researcher
gives a "5," obviously the interrater reliability would be inconsistent. Interrater reliability
is dependent upon the ability of two or more individuals to be consistent. Training,
education and monitoring skills can enhance interrater reliability.
7.8 Validity
Validity refers to the degree to which a study accurately reflects or assesses the specific
concept that the researcher is attempting to measure. It is the degree to which results
obtained from the analysis of data actually represent the phenomenon under study. It is
the accuracy and meaningfulness of inferences, which are based on the research results. It
has to do with how accurately the data obtained in the study represents the variables of
the study. If such data is a true reflection of the variables, then inferences based on such
data will be accurate and meaningful. Validity is largely determined by the presence or
absence of systematic error in the data e.g. using a faulty scale to measure.
Construct validity can be broken down into two sub-categories: Convergent validity and
discriminate validity. Convergent validity is the actual general agreement among ratings,
gathered independently of one another, where measures should be theoretically related.
Discriminate validity is the lack of a relationship among measures which theoretically
should not be related.
To understand whether a piece of research has construct validity, three steps should be
followed. First, the theoretical relationships must be specified. Second, the empirical
relationships between the measures of the concepts must be examined. Third, the
49
empirical evidence must be interpreted in terms of how it clarifies the construct validity
of the particular measure being tested.
The usual procedure in assessing the content validity of a measure is to use professional
or experts in the particular field. The instrument is given to two groups of experts, one
group is requested to assess what concept the instrument is trying to measure. The other
group is asked to determine whether the set of items or checklist accurately represents the
concept under study.
50
concerning what was and wasn't measured) and (2) the extent to which the
designers of a study have taken into account alternative explanations for any
causal relationships they explore. In studies that do not explore causal
relationships, only the first of these definitions should be considered when
assessing internal validity. Internal validity depends on the degree to which
extraneous variables have been controlled for in the study
Internal and external validity are inversely related to each other.
51
LECTURE SEVEN EXERCISES
7.1 Define the term measurement
7.2 Describe the various levels of measurement
7.3 Explain the various sources of measurement differences
7.4 Differentiate between the various types of variables.
7.5 Differentiate between validity and reliability
7.6 Explain the various ways of assessing Reliability
7.7 Describe the various types of reliability clearly indicating how to measure each.
7.8 Explain the various threats to internal and external validity
52
LECTURE EIGHT: RESEARCH INSTRUMENTS
8 Lecture Objectives
8.1 QUESTIONNAIRES
Each item in the questionnaire is developed to address a specific objective, research
question or hypothesis of the study. The researcher must also know how information
obtained from each questionnaire item will be analysed.
53
Disadvantages of Unstructured or open – ended questions
There is a tendency of the respondents providing information, which does not answer
the stipulated research questions or objectives.
The responses given may be difficult to categorize and hence difficult to analyze
quantitatively
Responding to open ended questions is time consuming, which may put some
respondent off.
3 Contingency questions
In particular cases, certain questions are applicable to certain groups of respondents. In
such cases, follow-up questions are needed to get further information from the relevant
sub-group only. These subsequent questions, which are asked after the initial questions,
are called ‘contingency questions’ or ‘ filter questions’. The purpose of these kinds of
questions is to probe for more information. They also simplify the respondent’s task, in
that they will not be required to answer questions that are not relevant to them.
4 Matrix questions
These are questions, which share the same set of response categories. They are used
whenever scales like likert scale are being used.
54
10. Simple words that are easily understandable should be used.
11. Questions that assume facts with no evidence should be avoided.
12. Avoid psychologically threatening questions.
13. Include enough information in each item so that it is meaningful to the
respondent.
55
The letter of transmittal / Cover letter
The letter of transmittal / Cover letter should accompany every questionnaire.
Contents of a letter of transmittal
It should explain the purpose of the study.
It should explain the importance and significance of the stuidy.
A brief assurance of confidentiality should be included in the letter.
If the study is affiliated to a certain institution or organisation, it is advisable to have
an endorsement from such an institution or organisation.
In a sensitive research, it may be necessary to assure the anonymity of respondents.
The letter should contain specific deadline dates by which the completed
questionnaire is to be returned.
Follow-up techniques
Sending a follow-up letter which should be polite, and asking the subjects to
respond
A questionnaire and a follow-up letter.
Response rate
It refers to the percentage of subjects who respond to questionnaires. Many authors
believe that a response rate of 50% is adequate for analysis and reporting. If the response
rate is low, the researcher must question the representativeness of the sample.
8.2 INTERVIEWS
An interview is an oral (face to face) administration of a questionnaire or an interview
schedule. To obtain accurate information through interviews, a researcher needs to obtain
the maximum co-operation from respondents. Interviews are particularly useful for
getting the story behind a participant's experiences. The interviewer can pursue in-depth
information around a topic. Interviews may be useful as follow-up to certain respondents
to questionnaires, e.g., to further investigate their responses. Usually open-ended
questions are asked during interviews.
56
6. Tell them how to get in touch with you later if they want to.
7. Ask them if they have any questions before you both get started with the interview.
8. Don't count on your memory to recall their answers. Ask for permission to record the
interview or bring along someone to take notes.
Sequence of Questions
1. Get the respondents involved in the interview as soon as possible.
2. Before asking about controversial matters (such as feelings and conclusions), first
ask about some facts. With this approach, respondents can more easily engage in
the interview before warming up to more personal matters.
3. Intersperse fact-based questions throughout the interview to avoid long lists of
fact-based questions, which tends to leave respondents disengaged.
4. Ask questions about the present before questions about the past or future. It's
usually easier for them to talk about the present and then work into the past or
future.
5. The last questions might be to allow respondents to provide any other information
they prefer to add and their impressions of the interview.
Wording of Questions
Wording should be open-ended. Respondents should be able to choose their own
terms when answering questions.
Questions should be as neutral as possible. Avoid wording that might influence
answers, e.g., evocative, judgmental wording.
Questions should be asked one at a time.
Questions should be worded clearly. This includes knowing any terms particular to
the program or the respondents' culture.
Be careful asking "why" questions. This type of question infers a cause-effect
relationship that may not truly exist. These questions may also cause respondents
57
to feel defensive, e.g., that they have to justify their response, which may inhibit
their responses to this and future questions.
While Carrying Out Interview
Occasionally verify the tape recorder (if used) is working.
Ask one question at a time.
Attempt to remain as neutral as possible. That is, don't show strong emotional
reactions to their responses. Patton suggests to act as if "you've heard it all before."
Encourage responses with occasional nods of the head, "uh huh"s, etc.
Be careful about the appearance when note taking. That is, if you jump to take a
note, it may appear as if you're surprised or very pleased about an answer, which
may influence answers to future questions.
Provide transition between major topics, e.g., "we've been talking about (some
topic) and now I'd like to move on to (another topic)."
Don't lose control of the interview. This can occur when respondents stray to
another topic, take so long to answer a question that times begins to run out, or
even begin asking questions to the interviewer.
Personal interviews
People selected to be part of the sample are interviewed in person by a trained
interviewer.
Requirements for success
Three broad conditions must be met in order to have a successful personal interview:
The participant must possess the information being targeted by the investigative
questions
The participant must understand his or her role in the interview as the provider of
accurate information
The participant must perceive adequate motivation to cooperate
The technique of stimulating participants to answer more fully and relevantly is termed
probing. Since it presents a great potential for bias, a probe should be neutral and appear
58
as a natural part of the conversation. Appropriate probes should be specified by the
designer of the data collection instrument. There are several probing styles e.g.
A brief assertion of understanding and interest e.g. comments such as “I see” “yes”.
An expectant pause
Repeating the question
Repeating the participant’s reply
A neutral question or comment
Question clarification.
59
Illiterate and functionally illiterate respondents can be reached
Interviewer can prescreen respondent to ensure he / she fits the population profile.
Responses can be entered directly into a portable microcomputer to reduce error
and cost when using computer assisted personal interviewing.
Telephone interviews
People selected to be part of the sample are interviewed on the telephone by a trained
interviewer.
Advantages of Telephone interviews
Lower costs than personal interviews
Expanded geographic coverage without dramatic increase in costs
Uses fewer, more highly skilled interviewers
Reduced interview bias
Fates completion time
Better access to hard-to-reach respondents through repeated callbacks
Can use computerized random digit dialing
Responses can be entered directly into a computer file to reduce error and cost when
using computer assisted telephone interviewing.
60
Be relaxed and friendly.
Be very familiar with the questionnaire or the interview guide.
Have a guide which indicates what questions are to be asked and in what order.
Interact with the respondent as an equal.
Pretest the interview guide before using it to check for vocabulary, language level
and how well the questions will be understood.
Inform the respondent about the confidentiality of the information given.
Not ask leading questions
Remain neutral in an interview situation in order to be as objective as possible.
An interview schedule
It’s a set of questions that the interviewer asks when interviewing. It makes it possible
to obtain data required to meet specific objectives of the study.
Advantages
It facilitates data analysis since the information is readily accessible and already
classified into appropriate categories.
If taken well, no information is left out.
Tape recording
The interviewer’s questions and the respondent’s answers are recorded either using a tape
recorder or a video tape.
Advantages
It reduces the tendency for the interviewer to make unconscious selection of data in
the course of the recording.
The tape can be played back and studied more thoroughly.
A person other than the interviewer can evaluate and categorize responses.
It speeds up the interview.
Communication is not interrupted.
Disadvantages
It changes the interview situation since respondents get nervous.
Respondents may be reluctant to give sensitive information if they know they are
being taped.
Transcribing the tapes before analysis is time consuming and tedious.
61
Advantages of interviews
It provides in-depth data, which is not possible to get using a questionnaire.
It makes it possible to obtain data required to meet specific objectives of the study.
Are more flexible than questionnaires because the interviewer can adapt to the
situation and get as much information as possible.
Very sensitive and personal information can be extracted from the respondent.
The interviewer can clarify and elaborate the purpose of the research and effectively
convince respondents about the importance of the research.
They yield higher response rates
Disadvantages of interviews
They are expensive – traveling costs
It requires a higher level of skill
Interviewers need to be trained to avoid bias
Not appropriate for large samples
Responses may be influenced by the respondent’s reaction to the interviewer.
8.3 OBSERVATION
Observation is one of the few options available for studying records, mechanical
processes, small children and complex interactive processes. Data can be gathered as
the event occurs. Observation includes a variety of monitoring situations that cover non-
behavioural and behavioural activities.
Advantages of observation
Enables one to:
Secure information about people or activities that cannot be derived from
experiment or surveys
Reduces obtrusiveness
Avoid participant filtering and forgetfulness
Secure environmental context information
Optimize the naturalness of the research setting
62
Limitations of observation
Difficulty of waiting for long periods to capture the relevant phenomena
The expense of observer costs and equipment
Reliability of inferences from surface indicators
The problem of quantification and disproportionately large records
63
LECTURE EIGHT EXERCISES
8.1 Briefly explain three commonly used research instruments.
8.2 Distinguish between
i. Direct and indirect questions
ii. Open-ended and closed ended questions
8.3 What special problems do open-ended questions have? How can these be
minimized? In what situations are open-ended questions most useful?
8.4 Distinguish among response error, interviewer error and non-response error
8.5 How do environmental factors affect response rates in personal interviews? How
can we overcome these environmental problems?
8.6 Compare and contrast interviews, questionnaires and observation as tools for
collecting primary data.
8.7 What ethical risks are involved in observation?
64
LECTURE NINE: DATA ANALYSIS
9 Lecture Objectives
Data preparation
This includes editing, coding and data entry. These activities ensure the accuracy of the
data and their conversion from raw form to reduced and classified forms that are more
appropriate for analysis.
Editing
Editing detects errors and omissions, corrects them when possible and certifies that
minimum data quality standards have been achieved. The editor’s purpose is to
guarantee that data are:
Accurate
Consistent with intent of the question and other information in the survey
Uniformly entered
Complete
Arranged to simplify coding and tabulation
Field editing
In large projects, field editing review is a responsibility of the field supervisor. It should
be done soon after the data have been gathered. During the stress of data collection, the
researcher often uses ad hoc abbreviations and special symbols. Soon after the interview,
experiment or observation, the investigator should review the reporting forms. It is
difficult to complete what was abbreviated or written in shorthand or noted illegibly if the
entry is not caught that day. When entry gaps are present from interviews, a call back
should be made rather than guessing what the respondent ‘probably would have said’.
Self-interviewing has no place in quality research.
Central editing
For a small study, the use of a single editor produces maximum consistency. In large
studies, the tasks may be broken down so that each editor can deal with one entire
section. This approach will not identify inconsistencies between answers in different
65
sections. However, this problem can be handled by identifying points of possible
inconsistency and having one editor check specifically for them.
Coding
Coding involves assigning numbers or other symbols to answers so the responses can be
grouped into a limited number of classes or categories. The classifying of data into
limited categories sacrifices some data detail but is necessary for efficient analysis.
Coding helps the researcher to reduce several thousand replies to a few categories
containing the critical information needed for analysis. In coding, categories are the
partitioning of a set and categorization is the process of using rules to partition a body of
data.
Coding rules
The categories should be:
Appropriate to the research problem and purpose: Categories must provide the best
partitioning of data for testing hypotheses and showing relationships.
Exhaustive
Mutually exclusive
Derived from one classification principle
66
Content analysis follows a systematic process i.e.
Selection of a unitization scheme. The units may be syntactical, referential,
prepositional or thematic
Selection of a sampling plan
Development of recording and coding instructions
Data reduction
Inferences about the context
Statistical analysis
Content analysis guards against selective perception of the content, provides for the
rigorous application of reliability and validity criteria and is amenable to
computerization.
“Don’t know” replies
“Don’t know” replies are evaluated in light of the questions nature and the respondent.
While many don’t know are legitimate, some result from questions that are ambiguous or
from an interviewing situation that is not motivating. It is better to report don’t knows as
a separate category unless there are compelling reasons to treat them otherwise.
Data entry
Data entry converts information gathered by secondary or primary methods to a medium
for viewing and manipulation. Data entry is accomplished by keyboard entry from pre-
coded instruments, optical scanning, real time keyboarding, telephone pad data entry, bar
codes, voice recognition, optical mark recognition (OMR) and data transfers from
electronic notebooks and laptop computers. Database programs, spreadsheets and editors
in statistical software programs e.g. SPSS and SAS offer flexibility for entering,
manipulating and transferring data for analysis, warehousing and mining.
Data description
The objective of descriptive statistical analysis is to develop sufficient knowledge to
describe a body of data. This is accomplished by understanding the data levels for the
measurements we choose, their distributions and characteristics of location, spread and
shape. The discovery of miscoded values, missing data and other problems in the data set
is enhanced with descriptive statistics
There are three general areas that make up the field of statistics: descriptive statistics,
relational statistics, and inferential statistics:
67
9.2.1 Measures of Central Tendency
The Mean
The most commonly used measure of central tendency is the mean. To compute the
mean, you add up all the numbers and divide by how many numbers there are. It's not the
average nor a halfway point, but a kind of center that balances high numbers with low
numbers. For this reason, it's most often reported along with some simple measure of
dispersion, such as the range, which is expressed as the lowest and highest number. The
mean is used for interval data only.
The Median
The median is the exact midpoint in a ranked distribution of numbers. It's not the
average; it's the halfway point. There are always 50% of all observations above the
median and 50% below the median. In cases where there is an even set of numbers, you
average the two middle numbers. The median is best suited for ranked values of data that
are ordinal and interval.
The Mode
The mode is the most frequently occurring value. It's the closest thing to what people
mean when they say something is average or typical. The mode doesn't even have to be a
number. It will be a category when the data are nominal or qualitative. The mode can be
applied to nominal, ordinal, or interval data. It is always the value (either quantitative or
qualitative) that occurs the most often. The mode is useful when you have a highly
skewed set of numbers, mostly low or mostly high. You can also have two modes
(bimodal distribution) when one group of scores are mostly low and the other group is
mostly high, with few in the middle.
9.2.2 Measures of Dispersion
In data analysis, the purpose of statistically computing a measure of dispersion is to
discover the extent to which scores differ, cluster, or spread from around a measure of
central tendency. The most commonly used measure of dispersion is the standard
deviation. You first compute the variance, which is calculated by subtracting the mean
from each number, squaring it, and dividing the grand total (Sum of Squares) by how
many numbers there are. The square root of the variance is the standard deviation.
The standard deviation is important for many reasons. One reason is that, once you know
the standard deviation, you can standardize by it. Standardization is the process of
converting raw scores into what are called standard scores, which allow you to better
compare groups of different sizes. Standardization isn't required for data analysis, but it
becomes useful when you want to compare different subgroups in your sample, or
between groups in different studies. A standard score is called a Z-score (not to be
confused with a z-test), and is calculated by subtracting the mean from each and every
number and dividing by the standard deviation. Once you have converted your data into
standard scores, you can then use probability tables that exist for estimating the
likelihood that a certain raw score will appear in the population. This is an example of
using a descriptive statistic (standard deviation) for inferential purposes.
68
Frequency table arrays data from highest to lowest values with counts and
percentages. They are most useful for inspecting the range of responses and their
repeated occurrence.
Bar charts and pie charts are appropriate for relative comparisons of nominal data.
Histograms are optimally used with continuous variables where intervals group the
responses.
Stem and leaf displays present actual data values using a histogram type device
that allows inspection of spread and shape.
Box plots use the five-number summary to convey a detailed picture of a
distribution’s main body, tails and outliers.
Control charts displays sequential measurements of a process together with a
centre line and control limits. The selection of a control chart depends on the level
of data one is measuring. It helps manager’s focus on special causes of variation
by revealing whether a system is under control and substantiating results from
improvements.
The Pareto diagram is a bar chart whose percentages sum to 100 percent. The
causes of the problem under investigation are sorted in decreasing importance with
bar height descending from left to right. Its pictorial array reveals the highest
concentration of quality improvement potential in the fewest number of remedies.
(a) Correlation
The most commonly used relational statistic is the correlation coefficient, known as
Pearson's R. It is used to a measure the strength and direction of the relationship between
two variables. Interpretation of a correlation coefficient does not even allow the slightest
hint of causality. The most a researcher can say is that the variables share something in
common; that is, are related in some way. The more two things have something in
common, the more strongly they are related. There can also be negative relations, but the
important quality of correlation coefficients is not their sign, but their absolute value. A
correlation of -.58 is stronger than a correlation of .43, even though with the former, the
relationship is negative. The following table lists the interpretations for various
correlation coefficients:
Value Comment
0.8 to 1.0 Very strong
0.6 to 0.8 Strong
0.4 to 0.6 Moderate
0.2 to 0.4 Weak
0.0 to 0.2 Very weak
69
The most frequently used correlation coefficient in data analysis is the Pearson product
moment correlation. It is symbolized by the small letter r, and is fairly easy to compute
from raw scores using the following formula:
If you square the Pearson correlation coefficient, you get the coefficient of determination,
written as r squared It is the amount of variance accounted for in one variable by the
other. Large R can also be computed by using the statistical technique of regression, but
in that situation, it's interpreted as the amount of variance explained for one variable by
another.
(b) Regression
Regression is the closest thing to estimating causality in data analysis, and that's because
it predicts how much the numbers "fit" a projected straight line, known as linearity or
linear relationship. There are also advanced regression techniques for curvilinear
estimation. The most common form of regression, however, is linear regression, and the
least squares method to find an equation that best fits a line representing what is called
the regression of y on x. The procedure is similar to computing a calculus minima.
Instead of finding the perfect number, however, one is interested in finding the perfect
line, such that there is one and only one line (represented by equation) that perfectly
represents, or fits the data, regardless of how scattered the data points. The slope of the
line (equation) provides information about predicted directionality, and the estimated
coefficients (or beta weights) for x and y (independent and dependent variables) indicate
the power of the relationship. Use of a regression formula (not shown here because it's
too large; only the generic regression equation is shown) produces a number called R-
squared, which is a kind of conservative, yet powerful coefficient of determination.
Interpretation of R-squared is somewhat controversial, but generally uses the same
strength table as correlation coefficients, and at a minimum, researchers say it represents
"variance explained."
70
(c) Discriminant analysis
It is used to classify people or objects into groups based on several predictor variables.
The groups are defined by a categorical variable with two or more values, whereas the
predictors are metric. The effectiveness of the discriminant equation is based not only on
its statistical significance but also on its success in correctly classifying cases to groups.
(d) Conjoint analysis
It is a technique that typically handles non-metric independent variables. It allows the
researcher to determine the importance of product or service attributes and the levels or
features that are most desirable. Respondents provide preference data by ranking or rating
cards that describe products. These data become utility weights of product characteristics
by means of optimal scaling and log linear algorithms.
71
These refer to a variety of tests for inferential purposes. Z-tests are not to be confused
with z-scores. Z-tests come in a variety of forms, the most popular being: (1) to test the
significance of correlation coefficients; (2) to test for equivalence of sample proportions
to population proportions, as in whether the number of minorities you've got in your
sample is proportionate to the number in the population. Z-tests essentially check for
linearity and normality, allow some rudimentary hypothesis testing, and allow the ruling
out of Type I and Type II error.
F-tests are much more powerful, as they allow explanation of variance in one variable
accounted for by variance in another variable. In this sense, they are very much like the
coefficient of determination. One really needs a full-fledged statistics course to gain an
understanding of F-tests, so suffice it to say here that you find them most commonly with
regression and ANOVA techniques. F-tests require interpretation by using a table of
critical values.
T-tests are kind of like little F-tests, and similar to Z-tests. It's appropriate for smaller
samples, and relatively easy to interpret since any calculated t over 2.0 is, by rule of
thumb, significant. T-tests can be used for one sample, two samples, one tail, or two-
tailed. You use a two-tailed test if there's any possibility of bi-directionality.
(c) Chi-Square
A technique designed for less than interval level data is chi-square (pronounced kye-
square), and the most common forms of it are the chi-square test for contingency and the
chi-square test for independence. Other varieties exist, such as Cramer's V, Proportional
Reduction in Error (PRE) statistics, Yule's Q, and Phi. Essentially, all chi-square type
tests involve arranging the frequency distribution of the data in what is called a
contingency table of rows and columns. Marginals, which are estimates of error in
predicting concordant pairs in the rows and columns (based on the null hypothesis), are
then computed, subtracted from one another, and expressed in the form of a ratio, or
contingency coefficient. Predicted scores based on the null hypothesis are called expected
frequencies, and these are subtracted from observed frequencies (Observed minus
Expected). Chi-square tests are frequently seen in the literature, and can be easily done by
hand, or are run by computers automatically whenever a contingency table is asked for.
The chi-square test for contingency is interpreted as a strength of association measure,
while the chi-square test for independence (which requires two samples) is a
nonparametric test of significance that essentially rules out as much sampling error and
chance as possible.
72
The Mann-Whitney U test is similar to chi-square and the t-test, and used whenever you
have ranked ordinal level measurement. As a significance test, you need two samples,
and you rank (say, from 1 to 15) the scores in each group, looking at the number of ties.
A z-table is used to compare calculated and table values of U. The interpretation is
usually along the lines of some significant difference being due to the variables you've
selected.
The Kruskal-Wallis H test is similar to ANOVA and the F-test, and also uses ordinal,
multi-sample data. It's most commonly seen when raters are used to judge research
subjects or research content. Rank calculations are compared to a chi-square table, and
interpretation is usually along the lines that there are some significant differences, and
grounds for accepting research hypotheses.
It is assumed that the concepts of hypothesis testing were done in CMS 200: Business
Statistics, therefore, just as a reminder, here is a summary of the steps followed in
hypothesis, and the various test statistics used under various conditions.
HYPOTHESIS TESTING
Definitions
Hypothesis: It’s a statement about a population parameter developed for the purpose of
testing.
Hypothesis testing: It’s a procedure based on sample evidence and probability theory to
determine whether the hypothesis is a reasonable statement.
73
Identify the test statistic
A test statistic is the statistic that will be used to test the hypothesis e.g.
test statistic for testing hypothesis about is which has a student t distribution
74
We now have two different test statistic for testing the population mean. The choice of
which one to use depends on whether or not the population variance is known.
Chi-square notation
The value of such that the area to its right under the chi-square curve is equal to and
is denoted by . The value is the point such that the area to its right is .
Hence, the area to its left is .
75
Sampling distribution of the sample proportion
Sampling distribution of
The sample proportion is a approximately normally distributed, with mean and
standard deviation , provided that is large ( and ).
hypothesis of tests about mean and variance. The test statistic for is
is
76
Confidence interval estimator of when and are known is
The test statistic for when and are unknown and and is
where
The test statistic is student t distributed with degrees of freedom, provided that
the following conditions are satisfied:
The two population random variables ( ) are normally distributed
The two population variances are equal i.e.
The quantity is called the pooled variance estimate.
77
INFERENCE ABOUT THE DIFFERENCE BETWEEN TWO MEANS:
MATCHED PAIRS EXPERIMENT
The key to recognizing a matched pairs experiment is to watch for a natural pairing
between one observation in the first sample and one observation in the second sample. If
a natural pairing exists, the experiment involves matched pairs. In this case the samples
are not independent.
The test statistic for testing the matched pair’s mean difference is
and , the test statistic is where and are sample variances created by
The distribution
Variables that are distributed range from zero to infinity. The exact shape of the is
determined by two sets of degrees of freedom; degrees of freedom associated with
the numerator, which in this case is and degrees of freedom associated with
the denominator, which in this case is .
78
The test statistic is . However, since under the null hypothesis, we always test
ANALYSIS OF VARIANCE
Chi-square test of a multinomial experiment (Goodness-of-fit test)
A multinomial experiment is a generalized version of a binomial experiment that allows
for more than two possible outcomes on each trial of the experiment.
Test statistic
Rejection region
Rule of five
For the discrete distribution of the test statistic to be adequately approximated by the
continuous chi-square distribution, the conventional rule is to require that the expected
frequency for each cell be at least 5. Where necessary, cells should be combined in order
to satisfy this condition. The choice of cells to be combined should be made in such a
way that meaningful categories result from the combination.
79
CHI-SQUARE TEST OF A CONTIGENCY TABLE
A contingency table is a rectangular table which items from a population are classified
according to two characteristics. The objective is to analyze the relationship between two
qualitative variables i.e. to investigate whether a dependence relationship exists between
two variables or whether the variables are statistically independent. The number of
degrees of freedom for a contingency table with rows and columns is
.
Examples
1. A random sample of eight auto drivers insured with a company and having similar
auto insurance policies was selected. The following table lists their driving experience
(in years) and the monthly auto insurance premium (in Sh.000) paid by them.
Driving experience (Years) 5 2 12 9 15 6 25 16
Monthly auto insurance premium 64 87 50 71 44 56 42 69
(In Sh.000)
i. Find the least squares regression line by identifying the appropriate dependent
and independent variable
ii. Interpret the meaning of the constants calculated in part (i).
iii. Compute the coefficient of correlation and coefficient of determination and
interpret them.
Solution:
i.
80
ii. it indicates the rate at which the insurance premium reduces with an
additional year of driving experience
It indicates the amount of premium that would be paid by a driver
without any years of experience.
iii.
There is a strong negative relationship between the years of experience and the monthly
auto insurance premiums
Solution
There is
enough
evidence
to support the manufacturer’s claim. Customers of this manufacturer can
confidently assume that the standard deviation of the fills is less than 2 cc.
3. Until recently, most car salespeople have been men. However, in the past decade,
numerous automobile dealers have hired women as salespeople in the hope that they
will succeed in selling more cars to female customers. A researcher decided to
determine whether there is a difference in performance between female and male
salespersons. She took a random sample of eight saleswomen and recorded the
commission that each earned in the last year. Her original plan was simply to take a
random sample of eight salesmen and compare their commissions with those of he
sales women. However, she felt that, since years of sales experience would be a
81
factor, she should also determine the number of years of sales experience for each of
the eight women. The years of experience and last year’s commission are shown in
the following table.
The next step in the experiment involved finding eight salesmen who had the same
number of years of experience as the eight sales women. The years of experience and
last year’s commissions for these three salesmen are shown in the following table
Years of experience and last years commission for eight salesmen.
Years of experience 5 3 2 4 10 12 1 7
Commissions ($ 1,000s) 29 28 29 20 51 49 12 37
(a) Can we conclude at the 5% significance level that female salespeople perform
differently from male salespeople?
(b) Estimate the mean difference with 95% confidence.
Solution
(a) The researcher wants to compare the commissions of saleswomen and salesmen.
Because the two groups of salespeople were paired according to their years of
experience, we identify the experiment as a matched pairs experiment. It follows that
the parameter of interest is the mean paired differences.
To compute the value of the test statistic we have to find out the paired differences:
Years of COMMISSIONS Paired X2
experience Saleswomen Salesmen Differences (x)
5 35 29 6 36
3 27 28 -1 1
2 24 29 -5 25
4 22 20 2 4
10 55 51 4 16
12 52 49 3 9
1 14 12 2 4
7 44 37 7 49
82
From the paired differences, we compute
and
Therefore
Solution
Ho: P1= 0.45, P2 = 0.4, P3 = 0.15
HA: At least one of the is not equal to its specified vale.
Test statistic:
83
Rejection region :
Value of the test statistic: assuming that the null hypothesis is correct, we can
calculate the expected number of consumers who prefer A, B and others using the
formula .
Company Observed Expected
frequency frequency
Therefore
Conclusion: Reject Ho
There is sufficient evidence at the 5% level of significance to allow us to conclude
that the market shares have changed from the levels they were at before the
advertising campaigns occurred.
5. The operations manager at a shirt manufacturing plant has been concerned about the
large number of defects that the company’s three shifts have been producing. They
appear to be three types of defects: Improper stitching, buttons not aligned with
button holes and inconsistent colouring. The manager decides to investigate the
problem. As a first step to improving the quality, she wants to know if the number
and type of defects are the same for all three shifts. A random sample of one day’s
shirt production is taken. The number of each type of defect and the number of
perfect shirts for each are shown in the following table.
Shift
Shirt condition 1 2 3 Total
Perfect 224 249 238 711
Improperly stitched 15 19 21 55
Unaligned buttons 8 12 12 32
Inconsistent colour 17 16 11 44
Total 264 296 282 842
Do these results allow the operations manager to conclude that at the 10%
significance level, there are differences in quality among the three shifts?
Solution
We need to conduct a chi-square of the contingency table to determine whether the
classifications are statistically independent.
Ho: The two classifications are independent
HA: the two classifications are dependent
Test statistic:
Rejection region :
84
The value of the test statistic
To compute the expected values for each cell, multiply the row total by the column total
and divide by the total number of shirts sampled. The expected cell frequencies are
shown in the parentheses in the following table
Shift
Shirt condition 1 2 3 Total
Perfect 224(222.93) 249(249.95) 238(238.13) 711
Improperly stitched 15(17.24) 19(19.33) 21(18.42) 55
Unaligned buttons 8(10.03) 12(11.25) 12(10.72) 32
Inconsistent colour 17(13.80) 16(15.47) 11(14.74) 44
Total 264 296 282 842
85
LECTURE NINE EXERCISES
9.1 Discuss the various tools that a researcher can use to analyse data.
9.2 Explain the conditions under which the following can be used to analyse data:
(a) Descriptive statistics
(b) Relational statistics
(c) Inferential statistics
9.3 A manufacturer of a new cheaper type of light bulb claims that his product is just
as well made and just as reliable as the higher priced competitive light bulbs. The
average life of the other light bulbs is known to be 5000 hours. In order to
examine the manufacturer’s claim, 50 of his bulbs were left on until they burned
out. The average length of life in the sample is 5,100 hours. With = 0.05, is
there sufficient evidence to reject the manufacturer’s claim? (Assume that
hours.
9.4 The manager of a departmental store is thinking about establishing a new billing
system for the stores credit customers. After a thorough financial analysis, she
determines that the new system will not be cost effective if the average monthly
account is less than 70,000. A random sample of 200 monthly accounts is drawn,
for which the mean monthly account is Sh. 66,000. With = 0.05, is there
sufficient evidence to conclude that the new system will not be cost effective?
Assume that the population standard deviation is Sh. 30,000.
9.5 Despite some controversy, scientists generally agree that high fibre cereals reduce
various forms of cancer. However, one scientists claims that people who eat high
fibre cereal for breakfast will consume on average fewer calories for lunch than
people who do not eat high-fibre cereal for break fast. If this is true, high-fibre
cereal manufacturers will be able to claim another advantage of eating their
products- potential weight reduction for dieters. To test the claim, 200 people
were randomly sampled and asked what they regularly eat for breakfast and
lunch. Each person was identified as either a consumer or a non-consumer of high
fibre cereal, and the number of calories consumed at lunch was measured and
recorded. These data are summarized below:
Calories consumed at lunch
Consumer of high Non -Consumer of high
fibre cereal fibre cereal
9.6 The manager of a large production facility believes that worker productivity is a
function of among other things the design of the job, which refers to the sequence
86
of worker movements. Two designs are being considered for the production of a
new product. To help decide which should be used, an experiment was performed.
Six randomly selected workers assembled the product using design A and another
eight workers assembled the product utilizing design B. the assembly times are
normally distributed as shown below:
Design A: 8.3, 5.3, 6.5, 5.1, 9.7, and 10.8
Design B: 9.5, 8.3, 7.5, 10.9, 11.3, 9.3, 8.8, and 8.0
(a) Can the manager conclude at the 5% significance level that the assembly times
differ for the two designs?
(b) Estimate with 99% confidence, the difference in mean assembly times between
design A and design B.
9.7 The owner of a large fleet of taxis is trying to estimate his costs for next years
operations. One major cost is for fuel purchases. Because of the high cost of
gasoline, the owner has converted his taxis to operate on propane. He needs to
know what the average consumption will be, so he decides to take a random
sample of eight taxis and measures the miles per gallon achieved. The results are
as follows:
28.1, 33.6, 42.1, 37.5, 27.6, 36.8, 39.0, 29.4
Estimate with 95% confidence the mean propane mileage for all taxis in his fleet?
(Assume that the distribution of mileage is normal).
9.8 A courier service advertises that its average delivery time is less than six hours for
local deliveries. A random sample of the amount of time this courier takes to
deliver packages to an address across town produced the following times
(rounded to the nearest hour).
7, 3, 4, 6, 10, 5, 6, 4, 3, 8
i. Is there sufficient evidence to support the courier’s advertisement at the 5%
level of significance?
ii. Find the 99% confidence interval estimate of the mean delivery time.
iii. What assumption must be made in order to answer part (i) and (ii) above.
9.9 A company manufactures steel shafts for use in engines. One method of judging
inconsistencies in the production process is to determine the variance of the
lengths of the shafts. A random sample of 10 shafts produced the following
measurements of their lengths in centimeters
20.5, 19.8, 21.1, 20.2, 18.9, 19.6, 20.7, 20.1, 19.8, 19.0
Find a 90% confidence interval estimate for the population variance assuming that
the lengths of the steel shafts are normally distributed.
9.10 One important factor in inventory control is the variance of the daily demand for a
product. Management believes that demand is normally distributed with the
variance equal to 250. In an experiment to test this belief about , the daily
demand was recorded for 25 days. The data has been summarized as follows:
87
Do the data provide sufficient evidence to show that management’s belief about the
variance is untrue? (Use ).
9.11 Some traffic experts claim that the variability of automobile speeds is a critical
factor in determining how many accidents are likely to occur on a highway. The
greater the variability, the more the accidents. Suppose a random sample of 101
cars reveals that the mean and variance of their speeds are 57.3 Km/h and 88.7
(Km/h) 2 respectively.
i. Can we conclude at the 5% significance level that the variance of all cars
speeds exceeds 50 (Km/h) 2?
ii. Estimate the variance of all cars speeds with 99% confidence
9.12 In a test to compare the speeds of two types of equal size computers, eight long
programs written in Pascal were run on both computers. Then the amount of CPU
time in minutes was measured and recorded. The CPU times are normally
distributed.
PROGRAM CPU TIME (minutes)
Computer 1 Computer 2
1 28 32
2 52 47
3 103 110
4 15 12
5 72 75
6 49 55
7 62 72
8 26 30
(a) Can we conclude at the 10% significance level, that the average CPU time for
computer 1 is less than the average CPU time for computer 2?
(b) Estimate the mean difference in CPU time between computer 1 and computer 2.
9.13 Do waiters or waitresses earn larger tips? To answer this question, a restaurant
consultant undertook a preliminary study. The study involved measuring the
percentage of the total bill left as a tip for one randomly selected waiter and one
randomly selected waitress in each of eight large restaurants during a one week
period. The results are shown below.
Restaurant Percentage as a ratio of tip to total bill
Waiters Waitresses
1 12.3 13.1
2 9.8 10.7
3 14.2 13.3
4 7.5 8.0
5 10.2 11.0
6 11.5 11.4
7 15.6 14.8
8 12.1 13.5
88
(c) What conclusions can the consultant draw from these results? (Use )
(d) Estimate with 99% confidence the mean difference in tips between waiters and
waitresses.
9.14 A nationally known manufacturer of replacement shock absorber claims that that
its product lasts longer than the type of shock absorber that the car manufacturer
installs. To test this claim, eight cars each had one new original and one new
replacement shock absorber installed on the rear end and were driven until the
shock absorbers were no longer effective. In each case, the number of miles until
this happened was recorded; the results are shown below:
Number of miles (‘000s)
Car Original shock absorber Replacement shock absorber
1 42.5 43.8
2 37.2 41.3
3 50.0 49.7
4 43.9 45.7
5 53.6 52.5
6 32.5 36.8
7 46.5 47.0
8 39.3 40.7
Is there sufficient evidence at the 5% significance level to support the
manufacturer’s claim?
9.15 To determine if a single die, is balanced, or fair, the die was rolled 600 times. The
observed frequencies with which each of the six sides of the die turned up are
recorded in the following table: -
Face 1 2 3 4 5 6
Observed frequency 114 92 84 101 107 102
Is there sufficient evidence to conclude at the 5% level of significance, that the
die is not fair?
89
9.17 The trustee of a company’s pension plan has solicited the opinions of a sample of
the company’s employees regarding a proposed revision of the plan. A breakdown
of the responses is shown in the table below: -
Response Blue-collar White-collar Managers
workers workers
For 67 32 11
Against 63 18 9
90
LECTURE TEN: REPORT WRITING TECHNIQUES
10 Lecture objectives
10.1
By the end of the lecture, the students should be able to:-
Describe the various components of a research proposal
Describe the various components of a research report
Explain the characteristics of a good research proposal
Explain the characteristics of a good research report
Introduction
A quality presentation of research findings can have an inordinate effect on a reader’s or
a listener’s perceptions of a study’s quality. Recognition of this fact should prompt a
researcher to make a special effort to communicate skillfully and clearly. Research
reports contain findings, analysis, interpretations, conclusions and recommendations.
Research reports differ depending on their aims and their readership. Reports should be
clearly organized, physically inviting and easy to read. Writers can achieve these goals
if they are careful with mechanical details, writing style and comprehensibility.
The final research report will have what is contained in the proposal (apart from the time
schedule and budget) and in addition dedication, acknowledgement, chapter four: Data
analysis and findings and chapter five: Summary, conclusions and recommendations.
91
Table of contents and list of figures and tables
Any report with several sections that total more than six to ten pages should have a table
of contents. If there are many tables, charts or other exhibits, they should also be listed
after the table of contents in a separate list of tables or list of figures.
List of abbreviations and acronyms
All abbreviations and acronyms used in report should be explained. An abbreviation is a
short form of a word while an acronym is a contraction formed by taking the first letter of
several words.
Acknowledgements
During the research process, the researcher may require help from other individuals or
organisations. It would be necessary if the researcher acknowledged received from these
individuals and organisations.
Abstract
A proposal abstract is a summary of what the researcher intends to do. It should be brief,
precise and to the point.
10.2.2 Introduction
The introduction prepares the reader for the report by describing the parts of the report.
Background to the problem
In the background, the researcher should broadly introduce the topic under investigation.
The researcher introduces briefly the general area of study, and then narrows down to the
specific problem to be studied. The background enables the reader to have an idea of
what is happening regarding the area under investigation.
The problem Statement
The researcher states the problem under investigation. The researcher should describe the
factors that make the stated problem a critical issue to warrant the study. Relevant
literature can be referred to. It should be brief and precise.
The purpose of the study
It is a broad statement indicating what the researcher intends to do about the problem
being investigated.
The objectives of the study
Research objectives are those specific issues within the scope of the stated purpose that
the researcher wants to focus upon and examine in the study. The objectives should be
specific, measurable, achievable, reliable and time bound. Objectives guide the researcher
in formulating testable hypotheses.
Research questions
These are the questions, which the researcher would like to be answered by undertaking
the study. They should be formulated from the objectives of the study.
Hypothesis
A hypothesis is a researchers prediction regarding the outcome of the study. It states
possible differences, relationships or causes between two variables or concepts.
Hypothesis are derived from or based on existing theories, previous research, personal
observations or experiences. The test of a hypothesis involves collection and analysis of
data that may either support or fail to support the hypothesis. If the results fail to support
a stated hypothesis, it does not mean that the study has failed but it implies that the
92
existing theories or principles need to be revised or retested under various situations.
93
The rationale for the choice of analysis approaches should be clear. A brief commentary
on assumptions and appropriateness of use should be presented.
10.2.5 Data analysis and Findings
The objective is to explain the data rather than draw interpretations or conclusions. When
quantitative data can be presented, it should be done as simply as possible with charts,
graphics and tables. The data need not include everything collected. Only material
important to the reader’s understanding of the problem and the findings should be
included. Both findings that support or do not support the hypothesis should be included.
10.2.6 Summary and conclusions
The summary is a brief statement of the essential findings. Sectional summaries may be
used if there are many specific findings. These may be combined into an overall
summary. Conclusions represent inferences drawn from the findings. Conclusions may
be presented in a tabular form for easy reading and reference. Summary findings may be
subordinated under the related conclusion statement.
Recommendations
There are usually a few ideas about corrective actions. In academic research, the
recommendations are often further study suggestions that broaden or test understanding
of the subject area. In applied research, the recommendations will usually be for
managerial action rather than research action. The writer may offer several alternatives
with justifications.
References
The use of secondary data requires a reference or a bibliography. Proper citation, style
and formats are unique to the purpose of the report. The
Appendixes
The appendixes are the place for complex tables, statistical tests, supporting documents,
copies of forms and questionnaires, detailed descriptions of the methodology, instructions
to field workers and other evidence important for later support. The reader who wishes to
learn about technical aspects of the study and to look at statistical breakdowns will want a
complete appendix.
Time schedule
It is a listing of the major activities and the corresponding anticipated time period it will
take to accomplish that activity. The time is usually given in months. Activities to be
undertaken can always overlap.
Budget
A budget is a list of items that will be required to carry out the research and their
approximate cost. It should be detailed enough and precise on items needed, prices per
unit and total cost. Details of requirements in each budget will be governed by the type of
research.
10.3 Characteristics of a Good Proposal:
The need for the proposed activity is clearly established, preferably with data.
The most important ideas are highlighted and repeated in several places.
94
The objectives of the project are given in detail.
There is a detailed schedule of activities for the project, or at least sample portions
of such a complete project schedule.
Collaboration with all interested groups in planning of the proposed project is
evident in the proposal.
The commitment of all involved parties is evident, e.g., letters of commitment in
the appendix and cost sharing stated in both the narrative of the proposal and the
budget.
The budget and the proposal narrative are consistent.
The uses of money are clearly indicated in the proposal narrative as well as in the
budget.
All of the major matters indicated in the proposal guidelines are clearly addressed
in the proposal.
The agreement of all project staff and consultants to participate in the project was
acquired and is so indicated in the proposal.
All governmental procedures have been followed with regard to matters such as
civil rights compliance and protection of human subjects.
Appropriate detail is provided in all portions of the proposal.
All of the directions given in the proposal guidelines have been followed carefully.
Appendices have been used appropriately for detailed and lengthy materials which
the reviewers may not want to read but are useful as evidence of careful planning,
previous experience, etc.
The length is consistent with the proposal guidelines and/or funding agency
expectations.
The budget explanations provide an adequate basis for the figures used in building
the budget.
If appropriate, there is a clear statement of commitment to continue the project
after external funding ends.
The qualifications of project personnel are clearly communicated.
The writing style is clear and concise. It speaks to the reader, helping the reader
understand the problems and proposal. Summarizing statements and headings are
used to lead the reader.
95
Emphasize important material and de-emphasize secondary material through
sentence construction and judicious use of italising, underlining, capitalizing
and parentheses.
Use ample space and wide margins to create a positive psychological effect on
the reader.
Choose words carefully, opting for the known and short rather than the
unknown and long.
Repeat and summarize critical and difficult ideas so readers can have time to
absorb them.
Review the writing to ensure the tone is appropriate
Proof read the final document to correct any errors.
Use short paragraphs
Indent parts of text that represent listings, long quotations or examples.
Use headings and subheadings to divide the report and its major sections into
homogeneous topical parts.
96
REFERENCES
Bordens S. Kenneth and Abott B Bruce. (2006). Research Design and Methods: A
Process Approach. Tata Mc Graw Hill, New Delhi.
Cooper, R.D and Schindler, S.P. (2006). Business Research Methods. 9th edition. Tata
McGraw-Hill international, Singapore.
David Dooley. (2001). Social Research Methods. Prentice Hall of India , New Delhi.
Kothari, C.R. (2004). Research methodology: Methods and techniques. Revised 2nd
edition. New age international publishers, New Delhi.
Mugenda, M. O and Mugenda, G.A (2003). Research Methods: Quantitative and
Qualitative approaches. Laba – Graphics services, Nairobi.
Saunders M et al. (2007). Research Methods for Business Students. 3rd Edition. Pearson
education, South Asia.
97