Research Notes
Research Notes
M.A. (Education)
II - Semester
348 23
EDUCATIONAL RESEARCH
METHODOLOGY AND STATISTICS
IN EDUCATION
Authors
Dr Suman Lata, Lecturer, Ginni Devi Modi Girls (PG) College, Modinagar, Ghaziabad
Units (1.2, 2.3, 3.4-3.4.3, 6.2.2-6.2.3, 6.5, 10)
Dr Harish Kumar, Associate Professor, Amity Institute of Education, Amity University, Noida
Units (2.2, 3.2-3.3, 3.5-3.6, 4.2-4.3, 4.6, 5, 7, 8.3, 11.2.4, 11.2.6)
Dr Deepak Chawla, Distinguished Professor, Dean Research and Fellow Programme, International Management Institute
(IMI), New Delhi
Dr Neena Sondhi, Professor, International Management Institute (IMI), New Delhi
Units (2.4, 8.2, 9.2, 11.2.7, 12, 14.7-14.8)
JS Chandan, Professor, Medgar Evers College, City University of New York
Units (3.7-3.7.1, 6.3-6.4, 11.2.3)
Vikas® Publishing House: Units (1.0-1.1, 1.3, 1.4-1.10, 2.0-2.1, 2.5-2.9, 3.0-3.1, 3.4.4-3.4.5, 3.7.2, 3.8-3.12, 4.0-4.1, 4.4-4.5,
4.7-4.12, 6.0-6.2.1, 6.6-6.10, 8.0-8.1, 8.4-8.8, 9.0-9.1, 9.3-9.8, 11.0-11.2.2, 11.2.5, 11.3-11.7, 13, 14.0-14.2, 14.3-14.6, 14.9-14.13)
All rights reserved. No part of this publication which is material protected by this copyright notice
may be reproduced or transmitted or utilized or stored in any form or by any means now known or
hereinafter invented, electronic, digital or mechanical, including photocopying, scanning, recording
or by any information storage or retrieval system, without prior written permission from the Alagappa
University, Karaikudi, Tamil Nadu.
Information contained in this book has been published by VIKAS® Publishing House Pvt. Ltd. and has
been obtained by its Authors from sources believed to be reliable and are correct to the best of their
knowledge. However, the Alagappa University, Publisher and its Authors shall in no event be liable for
any errors, omissions or damages arising out of use of this information and specifically disclaim any
implied warranties or merchantability or fitness for any particular use.
Work Order No. AU/DDE/DE1-291/Preparation and Printing of Course Materials/2018 Dated 19.11.2018 Copies - 500
SYLLABI-BOOK MAPPING TABLE
Educational Research Methodology and Statistics in Education
Educational research refers to the systematic collection and analysis of data related
NOTES to the field of education. Research may involve a variety of methods. Research
may involve various aspects of education including student learning, teaching
methods, teacher training, and classroom dynamics.
Educational researchers generally agree that research should be rigorous
and systematic. However, there is less agreement about specific standards, criteria
and research procedures. Educational researchers may draw upon a variety
of disciplines. These disciplines include psychology, sociology, anthropology,
philosophy and statistics. Methods may be drawn from a range of disciplines.
Conclusions drawn from an individual research study may be limited by the
characteristics of the participants who were studied and the conditions under which
the study was conducted.
The book Educational Research Methodology and Statistics in Education
enables the students to develop an understanding about the measurement and
evaluation techniques used in education and in analysing the principles of test
construction, both educational and psychological. This will in turn prepare the
students, and future professionals, to become independent users of test information,
who can describe the problems in measurement, explain how to approach these
problems and solve them. The book also provides guidance on how to evaluate
and use information about specific tests and focuses on the basic issues of
measurement in education.
This book is written with the distance learning student in mind. It is presented
in a user-friendly format using a clear, lucid language. Each unit contains an
Introduction and a list of Objectives to prepare the student for what to expect in
the text. At the end of each unit are a Summary and a list of Key Words, to aid in
recollection of concepts learnt. All units contain Self-Assessment Questions and
Exercises, and strategically placed Check Your Progress questions so the student
can keep track of what has been discussed.
Self-Instructional
Material
Introduction to
BLOCK - I Educational Research
UNIT 1 INTRODUCTION TO
EDUCATIONAL RESEARCH
Structure
1.0 Introduction
1.1 Objectives
1.2 Areas of Educational Research and Problems Related to Teaching and
Learning Process
1.3 Research Problem: Selection of Problem, Defining the Problem,
Statement of the Problem
1.3.1 Defining the Research Problem
1.3.2 Problem Formulation
1.3.3 Statement of Research Problem
1.4 Review of Related Literature: Purpose of the Review, Identification of
the Related Literature, Organizing the Related Literature
1.5 Validity, Reliability and Norms
1.5.1 Validity: Meaning and Method
1.5.2 Reliability: Meaning and Methods
1.5.3 Norms
1.6 Answers to Check Your Progress Questions
1.7 Summary
1.8 Key Words
1.9 Self Assessment Questions and Exercises
1.10 Further Readings
1.0 INTRODUCTION
Research takes advantage of the knowledge which has accumulated in the past as
a result of constant human endeavour. It can never be undertaken in isolation of
the work that has already been done on the problems which are directly or indirectly
related to a study proposed by a researcher. A careful review of the research
journal, books, dissertations, theses and other sources of information on the problem
to be investigated is one of the important steps in the planning of any research
study. A review of the related literature must precede any well planned research
study.
The unit describes the specific purposes which are served by the review of
related literature. The unit provides a study guide to the researcher in identifying
related literature, and in locating, selecting and utilizing the primary and secondary
Self-Instructional
Material 1
Introduction to sources of information available in the library. The unit deals with procedure which
Educational Research
the researcher should adopt for organizing the related literature in a systematic
manner.
The unit also explains the significance of the reliability and validity in qualifying
NOTES
a test as good. It also examines the methods of computing reliability, factors affecting
reliability and validity, types of validity and factors affecting validity.
1.1 OBJECTIVES
Following are the areas of educational research that provide information about
problems related to teaching and learning process norms.
1. Educational psychology: It helps the teacher to understand the child in the
classroom in order to improve teaching learning process. These researches provide
the following information:
Identification of factors that encourage learning.
Understand the personality of children in the class.
Effects of parental and teacher’s attitude towards children on learning.
Understand the problems of physically and socially handicapped children
in school system.
Role of teachers and text books in removing delinquency in adults and so
on.
Role of physical/intellectual efficiencies and defects in learning.
Usefulness of learning theories in various educational setting.
Relative effectiveness of various learning theories by field experiment.
Relative effectiveness of socio-cultural forces on the development of children.
Self-Instructional
2 Material
2. Philosophy of education: It provides the following information: Introduction to
Educational Research
Role of logic in various areas of education from concept formation to theory
development.
Reorganisation of social structure and educational system. NOTES
Role of knowledge, beliefs and values in developing educational theories.
Finding new implication of ancient Indian philosophies in the present scenario.
Role of ideologies and religion for improving educational practice.
3. Sociology of education: this provide the following information:
Effects of changes in the demographic structure on education.
Effects of new education policy on the expansion of education and
employment
Role of educational institutions
Role of social and cultural factors in bringing about social and educational
equity.
Minority and their problems.
Reservation policy and its impact on social system.
4. Curriculum development:
Structure of curriculum in India from primary to higher level.
Analysis and organisation of curriculum in various subject.
Analysis of text books at different level of learning.
Curriculum in relation to needs of the learner and the society.
Modernisation of curriculum in relation to changing needs.
5. Comparative education:
Administrative and educational policies of different countries and their impact
on the society as a whole
Impact of economic progress on education
Comparison of educational progress in various countries of the world
Impact of various systems of education in the world on each other
6. Educational management and administration: This area can help us in the
following aspects:
Impact of educational planning and legislations on performance
Problems of educational system and its impact on performance
Role of teachers and principals in enhancing performance of students
Impact of recruitment policies on output
Supervision and performance
Liberalisation and privatisation of higher education. Self-Instructional
Material 3
Introduction to 7. Guidance and counselling: We can understand the following aspects in this
Educational Research
area of education research:
Identification of factors contributes in the success in the life of students.
NOTES Role of family and neighbourhood in making children adjust in the society.
Construction of tools for diagnosing adjustment problems of students.
8. Educational technology:
Development of new teaching strategies.
Role of technology in teaching learning process.
Application of psychology to solving teaching problems.
Application of technological equipments and law in education.
Development of new audio-visual aids.
9. Problems of Indian education:
Pre-primary education
Primary education
Secondary education
Higher education
Vocational and technical education
Non-formal education
Distance education
Recommendation of commissions and committees on education
Self-Instructional
4 Material
o Contribution to knowledge in the specific field Introduction to
Educational Research
o Required cooperation from the research guide
(ii) Define the research problem: Research problems can be related to either
the state of nature or to the relationship of variables. In defining the research NOTES
problem, the researcher should study the existing literature including books
and journals available in the field with an interdisciplinary perspective to
base his research topic on some reliable background. He should also
concentrate on the relevance of the present research with the past works.
(iii) Setting objectives: After selecting the topic and defining the research
problem, the researcher should mention the objective of research. This means
that he should explain what he aims to achieve through the research. His
objective should also include an explanation of the extent to which the
research work is related to the specific field.
(iv) Survey existing literature: To understand the basis of research, it is
important for the researcher to review existing literature. This involves:
o Surveying existing books available in the field
o Reviewing other published materials like articles, journals, reports and
conference proceedings
The researcher should then prepare his own index for a period, in
chronological order, in addition to his consultation of various indices.
(v) Determine the sample design: Often, we select only a few items for
universal study purposes, for example, blood testing on sample basis to
perform census inquiry. The item selected is technically known as a sample.
The researcher must decide the manner of selecting a sample or decide
about the sample design. A sample design is a definite plan determined for
data collection to obtain a sample from a given population. The various
types of sample designs are:
o Deliberate sampling
o Simple random sampling
o Systematic sampling
o Stratified sampling
o Quota sampling
o Cluster sampling
o Multi-stage sampling
o Sequential sampling
The researcher should decide the sample design after considering the nature
of inquiry and other related factors. Sometimes, several of these methods
of sampling are used in the same study, which in turn is called mixed sampling.
Self-Instructional
Material 5
Introduction to (vi) Collect the data: There are a variety of ways to collect data. Primary data
Educational Research
can be collected through experiments or through surveys. If the researcher
performs an experiment, he/she observes some quantitative measurements.
This helps him/her to examine the truth in his hypothesis. In the case of
NOTES survey, however, the researcher can adopt one or more of the following
ways to collect data:
o By observation
o Through personal interviews
o Through telephone interviews
o By mailing of questionnaires
o Through schedules
(vii) Execute the project: This is the most important step in the research
process. The researcher should ensure that the project is performed in a
logical way and on time. If a survey is to be carried out, steps need to be
taken to ensure that it is under statistical control so that the collected data is
in accordance with the predetermined standard of accuracy.
(viii) Analyse the data: After data collection, the researcher’s next task is to
analyse them. The bulk data should be compressed into a few manageable
groups and tables for further analysis. The researcher can analyse the
collected data by using various statistical methods.
(ix) Test the hypothesis: After the analysis of data, the researcher should test
the hypothesis, if any. He should check if the facts support the hypothesis
or are contrary to the hypothesis. Statisticians have developed tests like
Chi square test, T-test and F-test, for hypothesis testing. This testing further
results in either acceptance or rejection of the hypothesis.
(x) Generalize and interpret: The real value of research lies in its ability to
arrive at certain generalizations. If the researcher cannot find a hypothesis
to start with, he might seek to explain his findings on the basis of some
theory. This is called interpretation. This may give rise to new questions and
further lead to more research.
(xi) Prepare report or thesis: This is the concluding step of research, where
the researcher prepares the report of what has been done by him. Generally,
the report should be designed in accordance with the following layout:
o Preliminary pages: Here, the title, date, acknowledgements and
foreword with the table of contents should be mentioned.
o Main text: This should be divided into introduction, summary, main
report and conclusion.
o End matter: This should contain appendices, bibliography and index.
Self-Instructional
6 Material
A report needs to be written in simple language with a precise and objective Introduction to
Educational Research
style. Charts and illustrations should be included to lay emphasis on the
study of research.
1.3.1 Defining the Research Problem NOTES
Problem discovery puts the research process into action and identification of the
problem is the first step towards its solution. Properly and completely defining a
business problem is easier said than done. Actually, the research task may be to
define or evaluate an opportunity or to clarify a problem. The definition and
discovery of the research problem is viewed under this broader context. In research,
often, only symptoms are apparent to begin with. The adage ‘a problem well
defined is a problem half solved’ is worth remembering. The investigation gets a
sense of direction with an orderly definition of the research problem. A careful
attention to the problem definition allows a researcher to set the proper research
objectives. When the purpose of research is clear, the chances of collecting the
relevant and necessary information are greater.
1.3.2 Problem Formulation
Just because a problem has been discovered or an opportunity has been recognized
does not mean that the problem has been defined. A problem definition indicates a
specific managerial decision area to be clarified or a particular problem to be
solved. It specifies research questions to be answered and the objectives of the
research. The research problem formulation involves the following interrelated
steps:
Ascertaining the objectives of the decision-maker
Understanding the problem’s background
Identifying and isolating the problem, rather than its symptoms
Determining the unit of analysis
Determining the relevant variables
Stating the research objectives and research questions (hypotheses)
The above-mentioned process ensures that the real research objectives/
questions are identified for the proposed research.
1.3.3 Statement of Research Problem
Both the decision-makers and researchers expect that the problem definition efforts
should result in a statement of the research problem or research objectives. On
completion of the exercise of formulating the research problem, the researcher
must prepare a written statement(s) that clarifies any ambiguity about what s/he
hopes the research will accomplish. Writing a series of research questions and
Self-Instructional
Material 7
Introduction to hypotheses can add clarity to the statement of the business problem. These research
Educational Research
questions are the researcher’s translation of the business problem into a specific
need for inquiry; and hypothesis is an unproven proposition that tentatively explains
certain facts or phenomena, a proposition that is empirically testable. In other
NOTES
words, research objectives/hypotheses explain the purpose of research in
measurable terms and define standards of what the research should accomplish.
Review of the related literature; besides, allowing the researcher to acquaint himself/
herself with current knowledge in the field or area in which he/she is going to
conduct his/her research, serves the following specific purposes:
1. The review of related literature enables the researcher to define the limits of
his/her field. It helps the researcher to delimit and define his/her problem.
To use an analogy given by D Ary et al., (1972, p. 56) a researcher might
say:
The work of A, B and C has discovered this much about my question; the
investigations of D have added this much to our knowledge. I propose to
go beyond D’s work in the following manner.
The knowledge of related literature, brings the researcher up-to-date on
the work which others have done, and thus to state the objective clearly
and concisely.
2. By reviewing the related literature, the researcher can avoid unfruitful and
useless problem areas. He/She can select those areas in which positive
findings are very likely to result and his/her endeavours would be likely to
add to the knowledge in a meaningful way.
3. Through the review of related literature, the researcher can avoid unintentional
duplication of well established findings. It is no use to replicate a study
when the stability and validity of its results have been clearly established.
4. The review of related literature gives the researcher an understanding of the
research methodology which refers to the way the study is to be conducted.
Self-Instructional
It helps the researcher to know about the tools and instruments which proved
8 Material
to be useful and promising in the previous studies. The advantage of the Introduction to
Educational Research
related literature is also to provide insight into the statistical methods through
which validity of results is to be established.
5. The final and important specific reason for reviewing the related literature is
NOTES
to know about the recommendations of previous researchers listed in their
studies for further research.
Identifying the Related Literature
The first step in reviewing the related literature is identifying the material that is to
be read and evaluated. The identification can be made through the use of primary
and secondary sources available in the library.
In the primary sources of information, the author reports his/her own work
directly in the form of research articles, books, monographs, dissertations or theses.
Such sources provide more information about a study than can be found elsewhere.
Primary sources give the researcher a basis on which to make his/her own judgment
of the study. Though consulting such sources is a time consuming process for a
researcher, yet they provide a good source of information on the research methods
used.
In secondary sources of information, the author compiles and summarizes
the findings of the work done by others and gives interpretation of these findings.
In them, the author usually attempts to cover all of the important studies in an area
reported in encyclopedia of education, education indexes, abstracts, bibliographies,
bibliographical references and quotation sources. Working with secondary sources
is not time-consuming because of the amount of reading required. The disadvantage
of the secondary sources, however, is that the reader is depending upon someone
else’s judgments on the importance of the study.
The decision concerning the use of primary or secondary sources depends
largely on the nature of the research study proposed by the researcher. If it is a
study in an area in which much research has been reported, a review of the primary
sources would be a logical first step. On the other hand, if the study is in an area in
which little or no research has been conducted, a check of the secondary sources
is more logical. Sources of information, whether primary or secondary, are found
in a library. The researcher must, therefore, develop the expertise to use resources
without much loss of time and energy. To aid the researcher in locating, selecting
and utilizing the resources, a study guide is provided in relation to their use in
educational research.
A researcher should be familiar with the library, its facilities and services.
He/She should also be acquainted with the regulations governing the use and
circulation of materials. Many libraries use a printed guide that contains helpful
information. The guide uses a diagram to indicate the location of the stacks, the
periodicals section, reference section, reading rooms, and special collections of
books, microfilm or microcard equipment, manuscripts, or pamphlets. The guide
Self-Instructional
Material 9
Introduction to lists the periodicals to which the library subscribes and the names of special indexes,
Educational Research
abstracts, and other reference materials.
The regulations concerning the use of stacks, the use of reserve books, the
procedures for securing reference materials held by the library or those that may
NOTES
be borrowed from another library are also included in the guide.
Research scholars and other readers are usually issued a library card giving
them access to the stacks. They may take the help of library staff or may carry on
the independent searching for the books and other reference materials. After using
the books, it is desirable for the readers to leave them on the tables so that the
library staff will return them to their proper position on the shelves.
Sometimes a reference is not available in the library. In such a situation, the reader
must consult the ‘union’ catalog, which lists references found in other libraries.
Such references may be obtained in the following ways:
(i) By inter-library loan system: The reader requests the librarian to borrow
the desired reference from other library where it is available.
(ii) By requesting a pholostatic copy: The reader may request the librarian
to obtain the photostat of a page or a number of pages of a desired reference
from the source.
(iii) By requesting an abstract or translation of the portion of a desired
reference: Some large libraries have abstracting and translating service
that provides abstracts, or copied or translated portions of needed materials
at an established fee.
(iv) By requesting microfilm or microfiche: The reader may purchase a
microfilm that can be projected on library microfilm equipment. A microfiche
is a sheet of film that contains microimages of a printed manuscript or book.
Its development has been one of the most significant contributions to library
and information services by providing economy and convenience of storing
and distribution of long runs of scholarly materials.
An even more significant development is the ultra-fiche. It has the capacity
of 3,200 pages per fiche (Mittal, 1979, p. 10).
Various types of cameras are used to record microimage on roll film. Some
of which are described as follows:
Planetary Camera is either 35 mm or 16 mm still camera which is mounted
on a vertical column that can be moved up and down as per the requirement. At a
time, it can be loaded with 100 feet of roll film. It does not cost much.
Step-and-repeat camera is a costly camera which is specifically used to
automatically record microimages on microfiche one-by-one.
Rotary camera, like the planetary camera, records microimages on roll
film and has the capacity to change the reduction ratio as desired.
Self-Instructional
10 Material
Flow camera costs less than half of a planetary camera. Its reduction ratio Introduction to
Educational Research
is fixed unlike previously mentioned types of camera.
All these cameras make use of silver, diazo or vesicular film for recording
microimages.
NOTES
Generally, six types of Readers are used for reading microfilms or microfiche
(Mittal, 1979, p. 13).
(i) Cuddly Microfiche Reader is a portable reader and can be used by
keeping it in one’s lap. It is very cheap and can be lent to library
members for home use.
(ii) Microfilm and Microfiche Readers is a reader/printer machine that
can make copies from both microfilm and microfiche.
(iii) Universal Machines essentially archieve by reading the description,
storing it and printing it. The example is Universal Tuning Machine
(UTM) or computer.
(iv) Reader/Printer is a push-button machine which not only helps in
reading a microfilm/microfiche but is capable of producing a full-sized
paper copy of the frame on the screen.
(v) Production Printer/Enlarge Printer is an automatic machine and
can print the requisite number of copies of a microfilm or selected
portions of a microfilm. It is used for mass production of full-sized
copies of microfilms.
(vi) Xerox Copyflow Machine is a costly machine, and therefore, is
beyond the reach of ordinary libraries. In can print out a microfilm into
a readable size, and as such, a single copy of any requisite document
can be had at low cost and in less time.
The card catalog is the index to the entire library collection. It lists the details
of publications found in the library with the exception of serially published
periodicals.
Generally, the card catalog contains author, title, and subject cards arranged
alphabetically. A great deal of information about a book can be found on the
cards. Besides the title of the book and the name of the author, the reader will find
the date of birth of the author, the edition, the publication date, the number of
pages, and the name and location of the publisher. Other items listed on the cards
are bibliographies, maps, portraits, illustrations, tables, series (if any) in which a
book appears, a brief description of the book—whether the book is a translation
and who did the translation.
Library classification systems provide ingenious ways of systematizing the
placement and location of books. Every system is based upon a methodology that
is logical and orderly to the smallest detail. The two principal systems of library
classification in the US are the ‘Dewey Decimal’ system and the ‘Library of
Congress’ system.
Self-Instructional
Material 11
Introduction to The ‘Dewey Decimal’ system is a decimal plan with the numbers running
Educational Research
from 001 to 999.99. The ‘Library of Congress’ system is particularly used in large
libraries. It provides for 20 main classes instead of the 10 of the ‘Dewey Decimal’
system. The system uses letters of alphabet for the principal headings and numerals
NOTES for further sub-grouping.
In a library, all books have a call number or letter that appears in the upper
left-hand corner of the author, subject or title card and on the back of the book.
These call numbers or letters are used to arrange the books serially on the library
shelves and within each classification, the books are arranged alphabetically by
author’s last name.
Identifying the best available sources pertaining to a problem and extracting the
essential information from them is of much importance to a researcher. For this,
he/she must develop some library searching techniques so as to save his/her time
and effort. Van Dalen (1973, p. 88) has suggested the following valuable guidelines
for a researcher:
1. Before using a library, familiarize yourself with its layout, facilities, services,
and regulations.
2. Learn how to use the microform (microfilm and microfitche) readers,
photocopies, and other mechanical aids.
3. Look in the stacks and in the periodical, reference, reserved book, and
rare book rooms, the materials that you will use frequently are placed.
4. Schedule your work session in a library when you will encounter the least
competition for resources and services.
5. Make out call slips for all or most of the books needed in one session.
6. Copy all information that the librarian needs to obtain each reference for
you, and before closing the periodical index or card catalog, recheck and
rectify any errors or omissions.
7. Arrange to spend a block of time in the library that is sufficient to accomplish
a specific task.
8. When little time is available, clear up questions that can be answered quickly
through the help of reference books that are readily available.
9. Before initiating search for materials in a library, write down questions that
cover precisely the information you wish to locate and group the questions
in accordance with the areas in the library where the answers may be found.
10. Compile a list of the present and any previous names of periodicals,
organizations, government agencies, research agencies, collectors of
statistics, libraries and museums with special collections, and outstanding
authorities in your field.
11. Keep a list of the best reference books, indexes, handbooks, historical
studies, and legal references in your area of specialization.
Self-Instructional
12 Material
12. Obtain copies of the best bibliographies and reprints of significant research Introduction to
Educational Research
studies for your files.
13. Note which periodicals regularly or occasionally print bibliographies, reviews
of literature or such other reference material and the issues in which they NOTES
appear.
There are a number of references that may be useful to a researcher in the
field of education. To facilitate the search for such material, a researcher may
consult the following carefully compiled volumes:
Constance M. Winchell, ed., A Guide to Reference Books, 8th edn.
(Chicago: American Library Association, 1967). This comprehensive work has
biennial supplements to bring the up-to-date information in a number of languages.
It describes and evaluates about 7,500 references and a section is devoted to
education.
Albert J. Walford, Guide to Reference Material. This is a two-volume
work which covers (1) Science and Technology (1966) and (2) Philosophy and
Psychology, Religion, Social Sciences, Geography, and History (1968).
Mary N. Barton and Marion V. Bell (1962), Reference Books: A Brief
Guide for Students and Other Users of the Library. This guide is helpful but
considerably shorter.
International Guide to Educational Documentation (1955-1960),
(UNESCO, 1963). This is a one-volume international guide to educational books,
pamphlets, periodicals, occasional papers, films and sound recordings.
Arvid Burke and Mary Burke, Documentation in Education. This guide
provides an excellent introduction to literature in the field of education.
The Standard Periodicals Directory, (New York: Oxbridge Publishing
Co., 1964-date). This is a directory of over 30,000 entries and covers every type
of periodical, with the exception of local newspapers. It is published every year
and covers about 200 classifications which are arranged by subject. An alphabetical
index is provided.
Christine L. Wyner, Guide to Reference Books for School Media Centres,
(Littleton, Colo: Libraries Unlimited, 1973). This guide includes 2575 entries with
evaluative comments on reference books and selection tools for use in educational
institutions. It is indexed by author, subject and title.
Encylopedias. These serve as a store house of information and usually
contain well-rounded discussion and selected bibliographies that are prepared by
specialists. Encyclopedias are arranged alphabetically by subject and for each
field of research, they present a critical evaluation and summary of the work that
has been done. In addition, these suggest the research needed in the field and also
provide a selective bibliography.
Self-Instructional
Material 13
Introduction to The following list provides a sample of encyclopedias that researchers in
Educational Research
the field of education might use:
A Cyclopedia of Education, Paul Monroe, ed., 5 vol., (New York:
Macmillan, 1911-13). It is edited by Paul Monroe with the assistance of
NOTES
departmental editors and more than 1,000 individual contributors. It provides
excellent bibliographies and is extremely useful for historical and biographical
purposes.
The Encyclopedia of Education, ed., Lee C. Deighton, (New York : The
Macmillan Company and The Free Press, 1971). The encyclopedia includes more
than 1,000 articles. It offers a view of the institutions and people, of the processes
and products, found in educational practice. The articles deal with history, theory,
research, philosophy, as well as with the structure and fabric of education.
Encyclopedia of Modern Education, Henry D. Rivlin and H. Schueller,
ed., (New York: Philosphical Library, 1943). This comprehensive work of about
200 authorities has been edited by Henry D. Rivlin and H. Schueller. It stresses
present day problems, trends, theories, and practices. The articles are accompanied
by brief bibliographies and there is a system of cross references.
Encyclopedia of Educational Research, Walter Scott Monroe, ed., rev.
edn., (New York: Macmillan, 1950). Monroe’s Encyclopedia of Educational
Research was prepared under the auspices of the American Educational Research
Association. It aims to present a critical evaluation, synthesis and interpretation of
research studies in the field of education. All the articles, arranged alphabetically,
are provided with bibliographies.
Encyclopedia of Educational Research, Chester Harris, ed., 3rd edn.,
(New York: Macmillan, 1960). Harri’s Encyclopedia of Educational Research is
also prepared under the auspices of the American Educational Research
Association. It is not merely a revision of earlier editions, but it is completely a
rewritten volume that has attempted to put into a new perspective.
Encyclopedia of Educational Research, Robert L. Ebel, ed., 4th edn.,
(New York: Macmillan, 1969). Ebel’s Encyclopedia of Educational Research
provides concise summaries of research and many references for further research.
The articles deal with persistent educational problems and continual educational
concerns.
Encyclopedia of Educational Research, Harold E. Mitzel, ed., 5th edn.,
(New York: The Free Press: A Division of Macmillan Publishing Co., Inc., 1982).
The contents of encyclopedia have been classified under 18 broad headings
alphabetically ranging from ‘Agencies and Institutions Related to Education,
Counselling, Medical, and Psychological Services; Curriculum Areas, etc., to
Teachers and Teaching’. The new concepts and topics, viz., ‘Computer-Based
Education’, ‘Drug Abuse Education’, ‘Equity Issues in Education’, ‘Ethnography’
and ‘Neurosciences’ are also included in this volume. These additions reflect recent
events and developments in the world to which education must attend.
Self-Instructional
14 Material
The International Encyclopedia of Education, Torsten Husen and Introduction to
Educational Research
T. Neville Postlethwaite, ed., (New York: Pergamon Press, 1985). This publication
is the first major attempt to present an up-to-date overview on educational
problems, practices and institutions all over the world. The information available in
this volume provides answers to three basic questions: What is the state of the art NOTES
in the various fields of education?, What scientifically sound and valid information
is available? and What further research is needed in various aspects of education?
The Encyclopedia of Comparative Education and National Systems of
Education, T. Neville Postlethwaite, ed., (New York: Oxford Press, 1988). This
encyclopedia is in two parts: the first part presents a series of articles about
comparative education; the second part provides description of 159 different
systems of education in various countries.
International Encyclopedia of the Social Sciences, (New York: Macmillan
Co., 1968). It was prepared under the direction of 10 learned societies. This
reference work covers topics in all of the social sciences.
Encyclopedia of Child Care and Guidance, (Garden City, New York:
Doubleday and Co., 1968). It is a comprehensive treatment of the nature of the
problems of childhood. It also suggests the methods of dealing with such problems.
Encyclopedia of Social Work, (New York: National Association of Social
Workers, 1965). This reference work presents extensive articles on all aspects of
social work.
Encyclopedia of Philosophy, (New York: McGraw-Hill Book Co. 1971).
This encyclopedia contains more than 7,000 articles written by more than 2,000
contributors in all areas of science and engineering.
Encyclopedia of Philosophy, (New York: Macmillan, Free Press 1967).
It is an authoritative and comprehensive reference work covering both Western
and Eastern thought—ancient, medieval and modern.
Encyclopedia of Indian Education, (New Delhi: NCERT, 2004). It
provides a comprehensive description of various concepts, themes and systems
pertaining to Indian education in ancient, medieval, pre-independence and post-
independence periods.
Dictionaries. They serve as constant guides to the researcher. A few known
dictionaries are detailed below:
Dictionary of Education, (New York: McGraw-Hill Book Co., 1973).
This dictionary covers 33,000 technical and professional terms. It also includes
educational terms used in various countries.
Comprehensive Dictionary of Psychological and Psycho-Analytical
Terms, (New York: David McKay Company). It contains more than 13,000 terms.
All these are defined in non-technical terms.
Dictionary of Sociology, Totowa, N.J., (Littlefield, Adams and Co.). In
this dictionary, sociological terms are defined in non-technical language. Self-Instructional
Material 15
Introduction to Roget’s International Thesaurus of Words and Phrases, (New York:
Educational Research
Crowell, Collier and Macmillan). A Thesaurus is the opposite of a dictionary. One
turns to the Thesaurus when one has an idea, but does not yet have appropriate
word to convey it. Thesaurus lists together the synonyms and antonyms of words.
NOTES A researcher should use this reference in conjunction with a good dictionary to
ensure precision of expression.
Yearbooks, Almanacs and Handbooks, A large amount of current
information on educational problems, thought and practices may be found in
yearbooks, almanacs and handbooks. Some yearbooks cover a new topic of
current interest each year and some others give more general reviews of events. A
list of some yearbooks, almanacs and handbooks is given as under:
The Handbook of Research on Teaching, N. L. Gage (ed.), (Chicago:
Rand McNally & Co., 1963). This handbook presents a comprehensive research
information on teaching with extensive bibliographies.
The Rand McNally Handbook of Education, Arthur W. Foshay (ed.),
(Chicago: Rand McNally & Co, 1963). It is a convenient source compilation of
the most important facts about education in the United States. This handbook
provides a quick-reference comparison of education in England, France and Russia.
Education Yearbook, (New York: Macmillan Co., 1972-date). This is an
annual publication. It includes statistical data on major educational issues and
movements with a comprehensive bibliography and reference guide.
Mental Measurement Yearbook, (Highland Park, New Jersey: Grayphon
Press, 1938-date). It is compiled by Oscar K. Buros and provides a comprehensive
summary on psychological measurement and standardized tests and inventories.
It is published every four years and includes reviews on all significant books on
measurement and excerpts from book reviews appearing in professional journals.
Indian Mental Measurement Hand Book: Intelligence and Aptitude
Tests, (New Delhi: National Council of Educational Research and Training
(NCERT), 1991). The Handbook is one of major efforts of National Library of
Educational and Psychological Tests (NLEPTs) published by NCERT to present
before the researchers, a review of the standardized tests, particularly in the areas
of ‘Intelligence’ and ‘Aptitude’. It makes available the organized information on
tests developed in India and the Indian adaptations or standardizations of foreign
tests. The information covers not only tests which are commercially available to
test users, and those available for restricted use, but also tests for which only
specimen sets are available. Test reviews have been included in this Handbook in
order to help the readers to evaluate the tests more critically.
The Student Psychologist’s Handbook: A Guide to Sources, (Cambridge,
Mass: Schenkman Publishing Co., 1969). This handbook describes the major
content areas of psychology with sources of information, methods of data collection,
and the use of reference materials.
Self-Instructional
16 Material
Data Processing Yearbook, (Detroit: Frank H. Gille, 1952-date). This Introduction to
Educational Research
yearbook is published irregularly and includes articles on equipment, techniques,
and developments in data processing. It also provides information about institutions
offering data processing and computer courses.
NOTES
United Nations Statistical Yearbook, (New York: United Nations, 1949-
date). This is an annual publication. It presents statistical data on population, trade,
finance, communication, health and education.
World Almanac-Book of Facts, (New York: Newspaper Enterprise
Association, 1968-date). This reference guide is published annually. It provides
up-to-date statistics and data concerning events, progress and conditions in social,
educational, political, religious, geographical, commercial, financial and economic
fields.
The Standard Education Almanac. It provides a record of facts and
statistics on virtually every aspect of education.
Directories and Bibliographies. Directories are used by a researcher to
locate the names and addresses of persons, periodicals, publishers or organizations
when he/she wants to obtain information, about financial assistance or research
material and equipments. Directories may help a researcher to find people or
organizations who have similar professional interests or who can answer his/her
queries or help to solve his/her problems.
A few important directories in the US and the UK are as follows:
Guide to American Educational Directories. It lists in one volume over
12,000 educational and allied directories. The directories are listed
alphabetically and are arranged under subject headings.
The Education Directory, (Washington: US Office of Education,
Superintendent of Documents, 1912-date). This directory is published
annually in five parts. It deals with names, educational agencies, officials,
institutions and other relevant data.
NEA Handbook for Local, State and National Associations, (Washington,
DC: National Education Association, 1945-date). This is an annual
publication and contains listings and comprehensive reports of state and
national officers of affiliated associations and departments.
Educator’s World, (Englewood, Colo.: Fisher Publishing Co., 1972-date).
This is an annual guide to more than 1,600 education associations,
publications, research and foundations.
National Faculty Directory, (Detroit: Gale Research Co., 1964-date).
This annual publication lists alphabetically the names and addresses of more
than 300,000 full-time and part-time faculty members and administrative
officials of colleges and universities in the US.
Encyclopedia of Associations, (Detroit: Gale Research Co., 1964-date).
This directory lists alphabetically more than 14,000 national associations of Self-Instructional
Material 17
Introduction to the US. It includes information on membership, addresses, names of
Educational Research
executive secretaries and statement of purpose of these associations.
Directory of Exceptional Children, (Boston: Porter Sargent Publishing
Co., 1962-date). This directory provides a description of schools, camps,
NOTES
homes, clinics, hospitals and services for the socially mal-adjusted, mentally
retarded or physically handicapped in the US.
Mental Health Directory, (Washington, D.C.: National Institute of Mental
Health, Government Printing Office, 1964-date). This annual publication
lists national, state and local mental health agencies in the US.
American Library Directory, (New York: R.R. Bowker Co., 1923-date).
This directory provides a binnaual guide to private, state, municipal,
institutional and collegiate libraries in the US and Canada. It includes
information on special collections, number of holdings, staff salaries, budgets
and affiliations.
Kelley, Thomas (ed.) Select Bibliographies of Adult Education in Great
Britain, (London: National Institute of Education, 1952). Blackwell, A.M. A List
of Researches in Educational Psychology Presented for Higher Degrees in
the Universities of the United Kingdom and the Irish Republic from 1918.
(London: Newnes Educational Publishing Co., 1950).
In India, a very few bibliographical guides to educational research on a
national basis have appeared. Bibliography of Doctorate Theses in Science
and Arts accepted by the Indian Universities for 1946–48 and 1948–50 was
published by the Inter-University Board of India. These are listed under the
respective universities with subject sub-headings including education.
The Index. A periodical index serves the same purpose as the index of a
book or the card file of a library. It identifies the source of the article or of the
book cited by listing the titles alphabetically, under author and the readers should
read all such directions before trying to locate the references.
A list of some important educational indexes is given below:
Education Index, (New York: H.W. Wilson Co., 1929-date). One valuable
and work saving guide created for educators is Education Index. It is published
monthly (September through June), cumulated annually and again every three years.
It indexes more than 250 educational periodicals, and many yearbooks, bulletins,
and monographs published in the US, Canada, and Great Britain. The material on
adult education, business education, curriculum, educational administration,
educational psychology, educational research, exceptional children, higher
education, guidance, health and physical education, international education, religious
education, secondary education and teacher education are included in this index.
Canadian Education Index, (Ottawa, Ontario: Canadian Council for
Educational Research, 1965-date). This index is issued quarterly and indexes
periodicals, books, pamphlets, and reports published in Canada.
Self-Instructional
18 Material
Current Index to Journals in Education, (New York: Macmillan Introduction to
Educational Research
Information, 1969-date). This index is published monthly and cumulated six monthly
and annually. It indexes about 20,000 articles each year from more than 700
education and education-related journals under author and subject headings.
NOTES
ERIC Educational Documents Index, (Washington, D.C.: National
Institute of Education, Government Printing Office, 1966-date). This index is
published annually. It is a guide to all research documents in the ‘Educational
Resources Information Centre’ or ERIC collection.
Index of Doctoral Dissertations International, (Ann Arbor, Mich.: Xerox
University Microfilms, 1956-date). Published as the issue 13 of Dissertation
Abstracts International each year, it consolidates into one list all dissertations
accepted by American, Canadian, and some European universities during the
academic year, as well as those available in microfilm.
International Guide to Educational Documentation, (Paris: UNESCO).
This guide is published every five years. It indexes annotated bibliographies covering
major publications, bibliographies and national directories written in English, French
and Spanish.
British Education Index. This index is compiled by the Librarians of
Institutes of Education, and it includes references to articles of educational interest
published during the period of four years. The index covers more than 50
periodicals.
Index to Selected British Educational Periodicals, (Leeds: Librarians of
Institutes of Education, 1945-date). This index is issued thrice per year and it
covers 41 educational periodicals excluding those on fundamental and adult
education.
Information about new ideas and developments often appear in periodicals
long before it appears in books. There are many periodicals in education and in
other closely-related areas that are the best sources for reports on recent research
studies. Such periodicals give much more up-to-date treatment to current questions
in education than books possibly can. They also publish articles of temporary,
local or limited interest that never appear in book form. The periodicals of proper
dates are the best sources for determining contemporary opinion and status, present
or past.
It has been estimated that there are about 2,100 journals that are specifically
related to the field of education. In all such journals, one may also find articles of
interest devoted to psychology, philosophy, sociology, and other subjects.
All those engaged in educational research should become acquainted with
certain educational periodicals, and they should also learn to use the indexes to
them. Knowledge about the editor of a periodical, the names of its contributors,
and the associations or institutions publishing it may serve as clues in judging the
merit of the periodicals.
Self-Instructional
Material 19
Introduction to Ulrich’s Periodicals Directory; A Classified Guide to a Selected List of
Educational Research
Current Periodicals, Foreign and Domestic, (New York: Bowker), provides a
comprehensive list of periodicals relating to education. In this directory, periodicals
are grouped in a subject classification and are alphabetically arranged. Each entry
NOTES includes title, sub-title, date of origin, frequency of publication, annual index,
cumulative indexes, and item characteristics of each periodical.
In India, many periodicals are published by some associations or institutions.
They provide a medium for dissemination of educational research and exchange
of experience among research workers, teachers, scholars and others interested
in educational research and related fields and professions.
Abstracts include brief summaries of the contents of the research study or
article. They serve as one of the most useful reference guides to the researcher
and keep him/her abreast of the work being done in his own field and also in the
related fields.
In America, the most useful of these references are the following:
The review of educational research: It gives an excellent overview of
the work that has been done in the field and about the recent developments. This
publication, between 1931 and 1969, reviewed about every three years each of
the given 11 major areas of education: (i) Administration; (ii) Curriculum;
(iii) Educational Measurement; (iv) Educational Psychology; (v) Educational
Sociology; (vi) Guidance and Counselling; (vii) Language Arts, Fine Arts, Natural
Sciences, and Mathematics; (viii) Research Methods; (ix) Special Programmes;
(x) Mental and Physical Development; and (xi) Teaching Personnel.
Since June 1970, the Review of Educational Research has pursued a
policy of publishing unsolicited reviews of research topics of the contributor’s
choice. The role played by this publication in the past has been assumed by the
Annual Review of Educational Research.
Research in Education (RIE): This represents the most comprehensive
publication of research materials in education today. RIE is published monthly
since 1966 by the Educational Resources Information Centre (ERIC) and indexed
annually. Each monthly issue of RIE is divided into three sections: (1) Document
Section; (2) Project Section; and (3) Accession Numbers Section.
Psychological abstracts: This useful reference is published by the American
Psychological Association since 1927. It is published bimonthly and contains
abstracts of articles appearing in over 530 journals, mostly educational periodicals.
The biannual issues (January–June, July–December) contain both author and
subject index.
Education abstracts: This is a publication of UNESCO, which began in
1949 and has been published monthly except in July and August. Each introductory
essay devoted to a particular aspect of education is followed by abstract of books
and documents selected from various countries dealing with the topic under
Self-Instructional consideration.
20 Material
In addition to the above periodicals, a researcher may also consult the Introduction to
Educational Research
following publications:
(i) Annual Review of Psychology (1950-date)
(ii) Child Development Abstracts and Bibliography (1927-date) NOTES
(iii) Psychological Bulletin (1904-date)
(iv) Sociological Abstracts (1952-date)
(v) Educational Administration Abstracts (1966-date)
(vi) Sociology of Education Abstracts (1965-date)
(vii) Mental Retardation Abstracts (1964-date)
(viii) Dissertation Abstracts International (1952-date)
In India, National Council of Educational Research and Training (NCERT)
has been publishing Indian Educational Abstracts to serve the cause of educational
research through disseminating information about educational researches available
in public domain. The information contains abstracts of the researches carried out
in India and abroad relevant to Indian educational scene with bibliographic
information. This biannual periodical also includes abstracts of doctoral theses,
research projects, published researches in the form of books and articles in the
reputed journals.
Many professional periodicals and year books, in India and abroad, include
some reviews of research and technical discussions of educational problems in
one or all the issues of their series. A list of some of the publications are as follows:
USA: Journal of Educational Research, NEA Research Bulletin,
Educational and Psychological Measurement, Journal of Experimental
Education, Research Quarterly, Journal of Research in Music Education,
American Educational Research Journal, Reading Research Quarterly, Journal
of Educational Psychology, Journal of Psychology, Journal of Social
Psychology, Journal of Applied Psychology, Sociology of Education,
American Journal of Sociology, American Sociological Review, Sociology
and Social Research, Harvard Educational Review, Journal of Teacher
Education, Elementary School Journal, History of Education Quarterly, and
Educational Forum.
UK: British Journal of Educational Psychology.
India: Indian Educational Review, Journal of Psychological Researches,
Indian Journal of Applied Psychology, Indian Journal of Experimental
Psychology, Journal of Education and Psychology, The Education Quarterly,
Perspectives in Education, Journal of Educational Planning and
Administration, University News, Journal of Higher Education, Indian
Journal of Education.
Theses and dissertations are usually preserved by the universities that award
the authors their doctoral and masters degrees. Sometimes these studies are
Self-Instructional
Material 21
Introduction to published in whole or in part in various educational periodicals or journals. Because
Educational Research
the reports of many research studies are never published, a check of the annual list
of theses and dissertations issued by various agencies is necessary for a thorough
coverage of the research literature.
NOTES
In the US, references of doctoral dissertations in all fields, including
education, can be found in sources compiled by various agencies. For the period
1912–1938, the Library of Congress issued the annual List of American Doctoral
Dissertations for published studies. The Association of Research Libraries
published the list of Doctoral Dissertations Accepted by American Universities
from 1933–1934 to 1954–1955. This service was continued by the Index to
American Doctoral Dissertations 1956–1963, which became the American
Doctoral Dissertations, 1963–64 to date. It lists all doctoral dissertations accepted
by the American and Canadian universities and other educational institutions.
Dissertation Abstracts International, May 1970, abstracts dissertations
in the humanities, social sciences, physical sciences and engineering. It is published
monthly. For each dissertation, there is a 600 word abstract that provides the
researcher enough information to satisfy his/her needs. If a researcher wants to
read a complete copy of a dissertation that is presented in Dissertation Abstracts
International, he/she can purchase a microfilm or xerox copy from the University
Microfilms. The reference number for placing an order and price are provided in
the abstract.
In India, only a few universities publish abstracts of dissertations and theses
that have been completed at the institution.
Kurukshetra University, Kurukshetra (Haryana) published Abstracts of
[Link]. Dissertations, Vol. I, 1966; Abstracts of [Link]. Dissertations, Vol. II,
1967; Abstracts of [Link]. Dissertations, Vol. III, 1968; Abstracts of [Link]
Dissertations, Vol. IV, 1969; Abstracts of [Link]. Dissertations, Vol. V, 1970;
Abstracts of [Link]. Dissertations and Ph.D. Theses, Vol. VI, 1973.
M.B. Buch (ed.) A Survey of Research in Education, (Centre of Advanced
Study in Education, Baroda: M.S. University, 1973). This publication contains all
the research studies in education completed in Indian universities up to 1972. The
break up of the studies in the said volume is 462 Ph.D. studies and 269 project
research. The abstracts of all the studies have been classified into 17 meaningful
areas of education. They are (i) Philosophy of Education, (ii) History of Education,
(iii) Sociology of Education, (iv) Economics of Education (v) Comparative
Education, (vi) Personality, Learning and Motivation, (vii) Guidance and
Counselling, (viii) Tests and Measurement, (ix) Curriculum, Methods, and
Textbooks, (x) Educational Technology, (xi) Correlates of Achievement,
(xii) Educational Evaluation and Examination, (xiii) Teaching and Teaching
Behaviour, (xiv) Teacher Education, (xv) Educational Administration, (xvi) Higher
Education, and (xvii) Non-Formal Education.
Self-Instructional
22 Material
M.B. Buch, ed., Second Survey of Research in Education (1972–1978) Introduction to
Educational Research
(Baroda: Society for Educational Research and Development, 1979). This
publication incorporates 839 research studies completed during the period 1972–
1978 and follows the same pattern of organization of 17 research areas as A
Survey of Research in Education (1973). The first chapter gives a broad NOTES
perspective of the place and function of research for educational development
including historical account of the development of educational research in India.
Each subsequent chapter includes a report based on the abstracts of research
studies giving the trend of research in the area, including the gaps and high-lighting
the research priorities as perceived by the author. The abstracts are arranged
alphabetically for each area and continuously numbered throughout the volume.
Each abstract contains the title of the study, the objective and/or hypotheses
examined, methodology including the sample, tools of research, the statistical
techniques used, and the findings. A special feature of this publication is the
incorporation of a large number of studies on educational problems completed in
the university departments of social sciences and humanities other than the
departments of education. The trend reports are based not on the research
completed during the period 1972–1978, but on the total research activities during
the period 1940–1978.
M.B. Buch, ed., Third Survey of Research in Education (1978–1983),
New Delhi: National Council of Educational Research and Training, 1987. The
publication comprises 20 chapters beginning with a comprehensive review for the
general trend of research in education in India based on a quantitative and qualitative
analysis of the studies. The trend reports in different areas of education have been
developed by eminent educationists on the basis of studies conducted during the
period of four decades, from 1943 to 1983. In all, 1481 research abstracts have
been presented after being classified under the 17 areas. Each research abstract
reports in brief the problem, objectives of the study, research techniques adopted,
and the findings and conclusions of the study. A special feature of the volume is the
chapter on ‘Research on Indian Education Abroad’, which presents a review of
192 doctoral dissertations submitted to American and British universities, covering
a period of around two decades. Another significant inclusion in the volume is the
chapter on ‘Priorities in Educational Research’. The volume also makes available
at one place a complete list of all researches in education conducted in India till
1983.
M.B. Buch, ed., Fourth Survey of Research in Education (1983–1988),
New Delhi: National Council of Educational Research and Training, 1991.
This publication, available in two volumes, covers researches in education
till 1988. It comprises 31 chapters beginning with a comprehensive review of the
general trend of research followed by trend reports in different areas of education
developed by eminent educationists on the basis of studies conducted during the
period of about four-and-a-half decades—1943 to 1988. In all, 1,652 research
Self-Instructional
Material 23
Introduction to abstracts have been presented after classification in 29 areas. The volume makes
Educational Research
available a complete list of all the 4,703 educational researches conducted in
India since 1943. The Fourth Survey has a new dimension. There is a chapter on
review of researches at the [Link] level in Indian Universities.
NOTES
Fifth Survey of Educational Research (1988–1992), New Delhi: National
Council of Educational Research and Training, 1997. This publication is also
available in two volumes and covers researches in education conducted during
1989–1992. It has dealt with all the areas of research which were covered in the
Fourth Survey with the addition of a chapter on researches in “Distance Education
and Open Learning”.
Sixth Survey of Educational Research (1993–2000), New Delhi: National
Council of Educational Research and Training, 2006. The first volume of this
publication was released in 2006 and the second volume is still awaited. The
researches in the areas of philosophy of education, teacher education, vocational
education, science education, distance education and open learning, women
education, guidance and counselling, physical education, health education and
sports, language teaching, inclusive education, educational technology and
population education conducted in India during the period 1993–2000 have been
reported in the first volume.
Many articles of particular interest to a researcher may be located through
pamphlets and newspapers. Current newspapers provide up-to-date information
on speeches, seminars, conferences, new trends, and a number of other topics.
Old newspapers, which preserve a record of past events, movements and ideas
are particularly useful in historical inquiries. Some libraries catalog pamphlets and
newspapers in their reference sections.
Government documents are a rich source of information. They include
statistical data, research studies, official reports, laws and other material that are
not always available elsewhere. These are available in national, regional, state as
well as local level government offices.
Monographs are also major sources of information on ongoing research. In
the US, universities and teachers’ colleges publish many research studies in education
in the form of monographs. A few examples of these are Supplementary
Educational Monographs, Educational Research Monographs, and Lincoln
School Monographs. In England too, various institutes of education publish
monographs from time to time. In India, only a limited number of monographs are
published by some universities and research organizations.
School Research Information Service (SRIS), Direct Access to Reference
Information (DATRIX), and Psychological Abstract Search and Retrieval Service
(PASAR) in the United States provide a number of computer-generated reference
sources that may save a great deal of time and effort of the researcher. SRIS
operated by Phi Delta Kappa (Bloomington, Indiana) provides a computer printout
of abstracts for a moderate fee. DATRIX, a development of the University
Self-Instructional
24 Material
Microfilms (Ann Arbor, Michigan) provides computerized retrieval for Dissertation Introduction to
Educational Research
Abstracts, from 1928 to date. The researcher can procure information on
Microfiche or Xerographic copy of the complete dissertation which he needs,
from University Microfilms, on payment. The PASAR furnishes printouts of
abstracts of psychological journal articles, monographs, reports, and parts of books NOTES
for a moderate fee.
Organizing the Related Literature
After making the comprehensive survey of the related literature, the next step for
the researcher is to organize the pertinent information in a systematic manner. It
should be done in such a way as to justify carrying out the study by showing what
is known and what remains to be investigated in the topic of concern. According
to Ary et al. (1972, p. 67):
The hypotheses provide a framework for organizing the related
literature. Like an explorer proposing an expedition, one maps out the known
territory and points the way to the unknown territory he proposes to explore.
If the study has several aspects, or is investigating more than a single
hypothesis, this is done separately for each facet of the study.
One should avoid the temptation to present the literature as a series of
abstracts. Rather, it should be presented in such a way as to lay a systematic
foundation for the study.
The organization of the related literature involves recording the essential
reference material and arranging it according to the proposed outline of the study.
Once pertinent information has been identified, the researcher should record
certain essential information for locating the material on 3 × 5 inch index card to
serve as a bibliography card. To make writing of the final report simpler, it is
desirable that the information recorded in the bibliography card should appear, in
content and style, exactly as it will appear in the final report.
The basic information in the bibliography card should include name of the
author with last name first; title of the book or article; name of the publication (for
articles); name of the publisher; date of publication; volume number, page numbers
and library call number (for books). If some of this information is not available, the
specified space should be left blank so that the missing information can be included
immediately upon locating the references.
After recording the essential information on the bibliography cards, it is
necessary to arrange the cards according to the location of the material in the
library. For example, the researcher may list together all cards pertaining to the
material located in the periodical section. Similarly, all the material located in the
reserve section may constitute another list, and so on. Then the researcher should
make a systematic review of the material located in a specific section of the library
and after reviewing each reference on the list, he/she should proceed to another
list.
Self-Instructional
Material 25
Introduction to All the information likely to be used in the final report should be recorded
Educational Research
on 4 × 6 inch card to serve as content card. The information to be recorded on
the content cards will depend on the source from which it is taken. If it is from a
primary source, it may include brief bibliographic information comprising author’s
NOTES last name, brief title of the report, specific page numbers on which information is
located; sentence statement of the problem; brief description of the study; statements
of findings or conclusions, or both; a card code as to the aspect of the research to
which the material most closely relates.
The information to be recorded from the secondary source is somewhat
different from the primary source. Turney and Robb (1971, p. 55) have given the
following suggestions for recording information from a secondary source:
1. Provide brief bibliographic information (as with a primary source).
2. Record on a single card only those statements that are related to the same
topic (if all the information cannot be placed on one card, continue statements
on another card and staple to the first card).
3. Paraphrase, in complete statements, the most relevant ideas. Record direct
quotations only if they are stated concisely and effectively, and if paraphrasing
might change the meaning.
4. Place a page number and a paragraph number after each separate statement
indicating its location in the reference in case you need to review it again.
5. Code the cards (probably in the upper-right hand corner) according to
topic(s) to which it most closely relates.
For the preparation of the report of the related literature, the researcher
should arrange the bibliographic and content cards according to the proposed
outline of the problem. This can be done with the help of card code.
The report of the related literature should begin with an introductory
paragraph describing the organization of the report. After the introduction, the
researcher should present the studies most relevant to each aspect of the proposed
problem outline. Studies with similar and contradictory results should be reported
side-by-side without using excessive space.
Test
1. What is the importance of survey of related literature in educational research?
Illustrate by taking a specific research problem as to how the survey of the
related literature can be helpful at various stages.
2. Describe the procedure which the researcher should adopt in identifying
related literature, and in locating, selecting and utilizing the primary and
secondary sources of information available in the library.
3. What library skills are required for a thorough survey of literature related to
a research topic in education?
Self-Instructional
26 Material
4. Name some important reference books with author’s names and some Introduction to
Educational Research
important educational journals you would like to consult in connection with
the problem you have selected for research.
5. Describe the procedure which the researcher should adopt in organizing
NOTES
the related literature in a systematic manner.
Self-Instructional
28 Material
Introduction to
Method Types of reliability Procedure of administration Educational Research
measure
Split-Half method Measure of internal Apply the test once. Score two
consistency equivalent halves of test (e.g.,
odd items and even items), NOTES
correct correlation between
halves to fit whole test by
Spearman-Brown formula to
measure reliability of the test.
Kuder-Richardson Measure of Internal Give test once, score total and
method consistency apply Kuder- Richardson
formula to know the degree of
reliability.
Inter-rater method Measure of consistency Use a set of student response
requiring judgmental scoring to
two or more raters and have
them in dependently score the
responses.
1.5.3 Norms
A level of performance for a particular group is represented by a norm. Getting a
raw score on any psychological test is meaningless unless having an additional
interpretive data. Hence, the score on psychological test are generally interpreted
by reference to norms that represent the test performance of the standardised
sample. Norms are formed by determining what parsons in a representative group
actually do on a test. In order to ascertain more precisely the individual’s exact
position with reference to the standardised sample, the raw score is converted
into some relative measure.
Self-Instructional
Material 29
Introduction to
Educational Research 1.6 ANSWERS TO CHECK YOUR PROGRESS
QUESTIONS
NOTES 1. A sample design is a definite plan determined for data collection to obtain a
sample from a given population.
2. The interrelated steps involved in formulating a research problem are as
follows:
Ascertain the objectives of the decision-maker
Understand the background of the problem
Identify and isolate the problem, rather than its symptoms
Determine the unit of analysis
Determine the relevant variables
3. The first step in reviewing the related literature is identifying the material that
is to be read and evaluated. The identification can be made through the use
of primary and secondary sources available in the library. In the primary
sources of information, the author reports his/her own work directly in the
form of research articles, books, monographs, dissertations or theses. In
secondary sources of information, the author compiles and summarizes the
findings of the work done by others and gives interpretation of these findings.
4. A microfiche is a sheet of film that contains microimages of a printed
manuscript or book. Its development has been one of the most significant
contributions to library and information services by providing economy and
convenience of storing and distribution of long runs of scholarly materials.
5. A card catalog is the index to the entire library library collection. It lists the
details of publications found in the library, with the exception of serially
published [Link], the card catalog contains author, title and
subject cards arranged alphabetically.
6. A bibliography card should include the basic information like the name of
the author with last name first; title of the book or article; name of the
publication (for articles); name of the publisher; date of publication; volume
number, page numbers and library call number (for books). If some of this
information is not available, the specified space should be left blank so that
the missing information can be included immediately upon locating the
references.
7. The characteristics of reliability are as follows:
It refers to the preciseness of a measuring instrument.
It is the coefficient of internal consistency and stability
Self-Instructional
30 Material
8. The validity of a test is determined by measuring the extent to which it Introduction to
Educational Research
matches with a given criterion.
Self-Instructional
Material 31
Introduction to of library classification in the US are the ‘Dewey Decimal’ system and the
Educational Research
‘Library of Congress’ system.
Encyclopedias serve as a store house of information, and usually contain
well-rounded discussion and selected bibliographies that are prepared by
NOTES
specialists. Encyclopedias are arranged alphabetically by subject, and for
each field of research, they present a critical evaluation and summary of the
work that has been done.
Abstracts include brief summaries of the contents of the research study or
article. They serve as one of the most useful reference guides to the researcher
and keep him/her abreast of the work being done in his own field and also
in the related fields.
The hypotheses provide a framework for organizing the related literature. If
the study has several aspects or is investigating more than a single hypothesis,
this is done separately for each facet of the study.
The organization of the related literature involves recording the essential
reference material and arranging it according to the proposed outline of the
study.
The basic information in the bibliography card should include name of the
author with last name first; title of the book or article; name of the publication
(for articles); name of the publisher; date of publication; volume number,
page numbers; and library call number (for books). If some of this information
is not available, the specified space should be left blank so that the missing
information can be included immediately upon locating the references.
Reliability refers to consistency of scores obtained by some individuals when
re-tested with the test on different sets of equivalent items or under other
variable examining conditions.
Validity of a test refers to its truthfulness; it refers to the extent to which a
test measures what it intends to measure. Standardization of a test requires
the important characteristic viz., validity.
Self-Instructional
32 Material
Introduction to
1.9 SELF ASSESSMENT QUESTIONS AND Educational Research
EXERCISES
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J. P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
Material 33
Variables
UNIT 2 VARIABLES
NOTES Structure
2.0 Introduction
2.1 Objectives
2.2 Meaning and Types of Variables
2.3 Moderator and Delineating Variables
2.4 Operationalizing Variables
2.5 Answers to Check Your Progress Questions
2.6 Summary
2.7 Key Words
2.8 Self Assessment Questions and Exercises
2.9 Further Readings
2.0 INTRODUCTION
The research problem requires identification of the key variables under the particular
study. To carry out an investigation, it becomes imperative to convert the concepts
into empirically testable and observable variables. A variable is generally a symbol
to which we assign numerals or values. The important types of variables and their
significance have been discussed in this unit. Along with this, operationalization of
variables has also been discussed in the unit.
2.1 OBJECTIVES
An Ideal Experiment
Now consider another version of this experiment wherein some of the other variables
differ across conditions. These are confounding variables (highlighted below) and
the experiment being conducted is not ideal. In this experiment, if the concentration
levels of subjects vary between the two conditions this may have been caused by
the independent variable, but it could also have been caused by one or more of
the confounding variables. For instance, if the subjects in the noisy environment
have lower concentration levels, is it because it was louder, too hot or because
they were tested in the afternoon? It is not possible to tell and therefore, this is
less than ideal.
Self-Instructional
36 Material
Variables
Variables Quiet Condition Noisy Condition
Noise Level (IV) Low High
IQ (EV) Average Average
Room temperature (EV) 68 degrees 82 degrees
NOTES
Sex of subjects (EV) 60 per cent F 60 per cent F
Task difficulty (EV) Moderate Moderate
Time of day (EV) Morning Afternoon
Etc. (EV) Same as noisy environment Same as quiet environ.
Etc. (EV) Same as noisy environment Same as quiet environ.
A Non-Ideal Experiment
Controlling the Confounding Variables
There are ways by which the extraneous variables may be controlled to ensure
that they do not become confounding variables. All people-related variables can
be controlled through the process of random assignment which will most likely
ensure that the subjects will be equally intelligent, outgoing, committed, etc. Random
assignment does not necessarily ensure that this is the case for every extraneous
variable in every experiment. However, when a sample is large, it works very well
and the researcher’s motives for using this method will never be questioned.
One of the way in which situation variables or task variables can be controlled
is basically by keeping them constant. For instance, in the noise-concentration
experiment above, we could adjust the thermostat and thereby keep the room
temperature constant and test all the subjects in the same room. We would, of
course, hold the difficulty of the tasks constant by giving all subjects in both
environments the same task. It is common practice for instructions to be written
or recorded and presented to each subject in exactly the same way.
At time, the researcher cannot hold a situation or task variable constant. In
these situations too, random assignment can be of great help. Consider a situation
where the same room is not available for testing the two groups and, in fact, one
group is tested on a Monday in Room 1 and the other group on a Tuesday in
Room 2. In this situation, we can use random assignment which can result in half
the Monday subjects in Condition A and the rest in Condition B, and the same for
the Tuesday subjects. Hence both conditions will have roughly the same percentage
of subjects tested in Room 1 and 2. On the other hand, consider what would
happen if we did not use random assignment and instead tested the Monday
subjects in Condition A and the Tuesday subjects in Condition B. In this situation,
we have two confounding variables. Subjects in Condition A were tested on different
days of the week and in different rooms from those in Condition B. Any difference
in the results could have been caused by one or more of the independent variable,
the day of the week, or the room.
Self-Instructional
Material 37
Variables In other words, confounding variables are those aspects of a study or sample
that might influence the dependent variable and whose effect may be confused
with the effects of the independent variable. Confounding variables are of two
types:
NOTES
(a) Intervening variables: These variables are hypothetical variables used to
explain casual links between other variables. In many types of behavioural
research, the relationship between independent and dependent variables is
not a simple one of stimulus to response. Certain variables that cannot be
controlled or measured directly may have an important effect on the outcome.
These modifying variables intervene between the cause and the effect. For
example, in a classroom language experiment, a researcher is interested in
determining the effect of immediate reinforcement on learning the parts of
speech. He suspects that certain factors or variables other than the one
being studied may be influencing the result, even though they cannot be
observed directly. These factors may be anxiety, fatigue or motivation. These
factors cannot be ignored. Rather they must be controlled as much as possible
through the use of appropriate design. For example, a variable (as memory)
whose effect occurs between the treatment in a psychological experiment
(as the presentation of a stimulus) and the outcome (as a response) is difficult
to anticipate or is unanticipated, and may confuse the results.
(b) Extraneous variables: These are variables that are not the subject of an
experiment but may have an impact on the results. Hence, extraneous
variables are uncontrolled and could significantly influence the results of a
study. Often we find that research conclusions need to be questioned further
because of the influence of extraneous variables. For instance, a popular
study was conducted to compare, the effectiveness of three methods of
social science teaching. Ongoing, regular classes were used, and the
researchers were not able to randomize or control the key variables as
teacher quality, enthusiasm or experience. Hence, the influence of these
variables could be mistaken for that of an independent variable.
For instance, in a study which attempts to measure the effect of temperature
in a classroom on students’ concentration levels, noise coming into the class through
doors or windows can influence the results and is therefore an extraneous variable.
This may be controlled by soundproofing the room, which illustrates how the
extraneous variable may be controlled in order to eliminate its influence on the
results of the test.
The following are the types of extraneous variables:
Subject variables pertain specifically to the people being studied. These
people’s characteristics such as age, gender, health status, mood,
background, etc., are likely to affect their actions.
Experimental variables pertain to the persons conducting the experiment.
Factors such as gender, racial bias, or language influence how a person
behaves.
Self-Instructional
38 Material
Situational variables represent the environment factors which were Variables
prevalent at the time when the study or research was conducted. These
include the temperature, humidity, lighting, and the time of day, and could
have a bearing on the outcome of the experiment.
NOTES
Continuous variable is one wherein, any value is possible within the range
of the limits of the variable. For instance, the variable ‘time taken to run
the marathon’ is continuous since it could take 2 hours 30 minutes or3
hours 15 minutes to run the marathon. On the other hand, the variable
‘number of days in a month that a worker came to office’ is not a
continuous variable since it is not possible to come to office on 14.32
days.
Discrete variable is one that does not take on all values within the limits
of the variable. For instance, the response to a five-point rating scale
must only have the specific values of 1, 2, 3, 4, or 5. It cannot have a
decimal value such as 3.6. Similarly this variable cannot be in the form
of 1.3 persons.
Quantitative variable is any variable that can be measured numerically
or on a quantitative scale, at an ordinal, interval or ratio scale. For
example, a person’s wages, the speed of a car, or the person’s waist
size are all quantitative variables.
Qualitative variables are also known as categorical variables. These
variables vary with no natural sense of ordering. They are therefore
measured on the quality or characteristic. For example, eye colour (black,
brown, or blue) is a qualitative variable, as are a person’s looks (pretty,
handsome, ugly, etc.). Qualitative variables may be converted to appear
numeric, but this conversion is meaningless and of no real value (as in
male = 1, female = 2).
Self-Instructional
Material 39
Variables that moderates the relationship between environmental awareness campaign and
intention to purchase green products by the public.
Delineation variables are those variables which will be accounted on without
relating them to anything in particular.
NOTES
The most important variable to be studied and analysed in research study is the
dependent variable (DV). The entire research process is involved in either describing
this variable or investigating the probable causes of the observed effect. Thus, this
in essence has to be reduced to a measurable and quantifiable variable. For example,
in the organic food study, the consumer’s purchase intentions and the retailers
stocking intentions as well as sales of organic food products in the domestic market,
could all serve as the dependent variable.
A financial researcher might be interested in investigating the Indian
consumers’ investment behaviour, post the recent financial slow down. In another
study, the HR head at Cognizant Technologies would like to study the organizational
commitment and turnover intentions of short and long tenure employees in the
company.
Hence, as can be seen from the above examples, it might be possible that in
the same study there might be more than one dependent variable.
Any variable that can be stated as influencing or impacting the dependent
variable is referred to as an independent variable (IV). More often than not, the
task of the research study is to establish the causality of the relationship between
the independent and the dependent variable(s). The proposed relations are then
tested through various research designs.
In the organic food study, the consumers’ attitude towards healthy lifestyle
could impact their organic purchase intention. Thus, attitude becomes the
independent and intention the dependent variable. Another researcher might want
to assess the impact of job autonomy and role stress on the organizational
commitment of the employees; here job autonomy and role stress are independent
variables.
Moderating variables are the ones that have a strong contingent effect on
the relationship between the independent and dependent variables. These variables
Self-Instructional
40 Material
have to be considered in the expected pattern of relationship as they modify the Variables
Self-Instructional
Material 41
Variables Besides the moderating and intervening variables, there might still exist a
number of extraneous variables (EVs) which could affect the defined relationship
but might have been excluded from the study. These would most often account for
the chance variations observed in the research investigation. For example,
NOTES a tyrannical boss; family pressures or nature of the industry could impact the flexi-
time impact, but since these would be applicable to individual cases, they might
not heavily impact the direction of the findings. However, in case the effect is
substantial, the researcher might try to block their effect by using an experimental
and a control group.
At this stage, we can clearly distinguish between the different kinds of
variables discussed above. An independent variable is the prime antecedent
condition which is qualified as explaining the variance in the dependent variable;
the intervening variable follows the occurrence of the independent variable and
may in turn impact the dependent variable; the moderating variable is a contributing
variable which might impact the defined relationship; the extraneous variables are
outside the domain of the study and responsible for chance variations, but in some
instances, their effect might need to be controlled.
Concepts and Operationalization of Concepts
Having identified and defined the variables under study, the next step requires
operationalizing the stated relationship in the form of a theoretical framework. This
is an outcome of the problem audit conducted prior to defining the research problem;
it can be best understood as a schema or network of the probable relationship
between the identified variables. Another advantage of the model is that it clearly
demonstrates the expected direction of the relationships between the concepts.
There is also an indication of whether the relationship would be positive or negative.
This step however is not mandatory as sometimes the objective of the
research is to explore the probable variables that might explain the observed
phenomena (DV) and the outcome of the study helps to theorize and propose a
conceptual model.
The theoretical framework, once formulated, is a powerful driving force
behind the research process and ought to be comprehensively developed. It requires
a thorough understanding of both theory and opinion.
Given below is a predictive model for turnover intentions developed to
explain the high rate of attrition amongst BPO professionals. Once validated, it is
of course possible to test it in different contexts and differing respondent population.
The Turnover Intention Model
The proposed model to predict turnover intention is specified as mentioned below:
TI = f (WE, OC, A, MS, TWE) ...(2.1)
Where, TI = Turnover intention
Self-Instructional
WE = Work exhaustion
42 Material
OC = Organizational commitment Variables
A = Age
MS = Marital status
TWE = Total work experience NOTES
The theoretical construct of work exhaustion is influenced by Perceived
Workload (PWL), Fairness of Reward (FOR), Job Autonomy (JA) and Work
Family Conflict (WFC) [Adapted from Ahuja, Chudoba and Kacman, 2007] .
This can be mathematically written as:
WE = f (PWL, FOR, JA, WFC) ...(2.2)
Similarly, Organizational Commitment depends upon Job Autonomy,
Work–Family Conflict, Fairness of Reward and Work Exhaustion (WE)[Adapted
from—Ahuja, Chudoba and Kacman, 2007]. Therefore, this can be stated
mathematically as:
OC = f (JA, WFC, FOR, WE) ...(2.3)
The model is diagrammatically represented in Figure 2.1.
Self-Instructional
44 Material
being studied. According to Kerlinger, ‘variable is a property that takes on Variables
different value’.
2. Dependent variables represent characteristics that alter, appear or vanish
as a consequence of introduction, change or removal of independent
NOTES
variables. The dependent variable may be a test score or achievement of a
student in a test, the number of errors or measured speed in performing a
task.
3. Intervening variables are termed as hypothetical variables which are used
to explain casual links between other variables. Certain variables that cannot
be controlled or measured directly may have an important effect on the
outcome. These modifying variables intervene between the cause and the
effect.
4. Moderator is a special type of independent variable. It is a third variable
that affects the direction or strength of relationship between the independent
and dependent variable by changing the effect of the main variables. It may
be qualitative or quantitative.
5. Any variable that can be stated as influencing or impacting the dependent
variable is referred to as an independent variable (IV). More often than not,
the task of the research study is to establish the causality of the relationship
between the independent and the dependent variable(s).
6. The theoretical framework is a powerful driving force behind the research
process and ought to be comprehensively developed. It requires a thorough
understanding of both theory and opinion.
2.6 SUMMARY
Self-Instructional
Material 45
Variables A confounding variable is one which is not the subject of the study but is
statistically related with the independent variable. Hence, changes in the
confounding variable track the changes in the independent variable.
There are ways by which the extraneous variables may be controlled to
NOTES
ensure that they do not become confounding variables. All people-related
variables can be controlled through the process of random assignment which
will most likely ensure that the subjects will be equally intelligent, outgoing,
committed, etc.
Intervening variables are hypothetical variables used to explain casual links
between other variables. In many types of behavioural research, the
relationship between independent and dependent variables is not a simple
one of stimulus to response. Certain variables that cannot be controlled or
measured directly may have an important effect on the outcome.
Extraneous variables are not the subject of an experiment but may have an
impact on the results. Hence, extraneous variables are uncontrolled and
could significantly influence the results of a study.
Discrete variable is one that does not take on all values within the limits of
the variable.
Quantitative variable is any variable that can be measured numerically or on
a quantitative scale, at an ordinal, interval or ratio scale.
Qualitative variables are also known as categorical variables. These variables
vary with no natural sense of ordering. They are therefore measured on the
quality or characteristic.
Moderator variable is a third variable that affects the direction or strength
of relationship between the independent and dependent variable by changing
the effect of the main variables.
The most important variable to be studied and analysed in research study is
the dependent variable (DV). The entire research process is involved in
either describing this variable or investigating the probable causes of the
observed effect.
Any variable that can be stated as influencing or impacting the dependent
variable is referred to as an independent variable (IV). More often than not,
the task of the research study is to establish the causality of the relationship
between the independent and the dependent variable(s). The proposed
relations are then tested through various research designs.
Moderating variables are the ones that have a strong contingent effect on
the relationship between the independent and dependent variables. These
variables have to be considered in the expected pattern of relationship as
they modify the direction as well as the magnitude of the independent–
dependent association.
Self-Instructional
46 Material
In the organic food research, the objectives and sub-objectives of the study Variables
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Self-Instructional
Material 47
Variables Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
NOTES
Guilford, J. P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
48 Material
Hypothesis
UNIT 3 HYPOTHESIS
Structure NOTES
3.0 Introduction
3.1 Objectives
3.2 Concept of Hypothesis
3.3 Sources of Hypothesis
3.4 Types of Hypothesis
3.4.1 Directional Research Hypothesis
3.4.2 Non-Directional Research Hypothesis
3.4.3 Question From Hypothesis
3.4.4 Null Hypothesis
3.4.5 Alternative Hypothesis
3.5 Formulating Hypothesis
3.6 Characteristics of a Good Hypothesis
3.7 Hypothesis Testing and Theory and Errors in Testing of Hypothesis
3.7.1 Procedure for Hypothesis Testing
3.7.2 Committing Errors: Type I and Type II
3.8 Answers to Check Your Progress Questions
3.9 Summary
3.10 Key Words
3.11 Self Assessment Questions and Exercises
3.12 Further Readings
3.0 INTRODUCTION
Research takes advantage of the knowledge which has accumulated in the past as
a result of constant human endeavour. It can never be undertaken in isolation of
the work that has already been done on the problems which are directly or indirectly
related to a study proposed by a researcher. A careful review of the research
journal, books, dissertations, theses and other sources of informations on the
problem to be investigated is one of the important steps in the planning of any
research study. A review of the related literature must precede any well planned
research study.
Hypothesis is an assumption or proposition whose testability is to be tested
on the basis of the compatibility of its implications with empirical evidence with
previous knowledge (Mouly, 1963). It is also a declarative statement in which the
investigator makes a prediction or a conjecture about the outcome of the
relationship. The conjecture or the prediction is not simply an ‘educated guess’;
rather it is typically based on past researches, which investigators gather as evidence
to advance the hypothesized relationship between variables.
In this unit, you will also learn about the concept of hypothesis testing. For
this, a hypothesis needs to be appropriate. Testing a hypothesis means verification
Self-Instructional
Material 49
Hypothesis of the hypothesis. This unit will describe the application of hypothesis testing in a
variety of cases, such as comparing two related terms and testing equality of
variance of two normal populations. A number of hypothesis tests, such as t-test
and z-test, facilitate the process of hypothesis testing. The unit will also describe
NOTES the statistical techniques dealing with hypothesis testing.
3.1 OBJECTIVES
Self-Instructional
50 Material
A hypothesis relates theory to observation and vice-versa. Hypotheses when Hypothesis
tested are either rejected or accepted, and help to infer the conclusion, which
helps in theory building. Being a specific statement of prediction, a hypothesis
describes in concrete (rather than theoretical) terms what you expect will happen
in your study. Not all studies have hypotheses. Sometimes a study is designed to NOTES
be exploratory. In such researches, no formal hypothesis is established, and it may
be the case that the actual objective of the study is to explore one or more specific
areas more thoroughly in order to develop specific hypotheses or predictions that
could be tested through research in the future. A single study could result in one or
several hypotheses.
Some definitions of hypothesis are:
According to Townsend, ‘Hypothesis is defined as suggested answer
to a problem’.
According to McGuigan, ‘A hypothesis is a testable statement of a
potential relationship between two or more variables’.
According to Uma Sekaran, ‘A hypothesis is defined as a logically
conjectured relationship between two or more variables in the form
of testable statement. These relationships are based on theoretical
framework formulated for the research problem. The hypotheses
are often statements about population parameters like expected
value and variance, for example a hypothesis might be that the
expected value of the height of 10-year-old boys in the Scottish
population is not different from that of 10-year-old girls.’
According to Kerlinger, ‘A good hypothesis is one which satisfies
the following criteria:
(i) Hypothesis should state the relationship between variables.
(ii) They must carry clear implications for testing the stated
relations.’
This means that (a) statements contain two or more variables which can be
measured, (b) they must state clearly how the two or more variables are related,
and (c) it is important to note that facts and variables are not tested but relations
between variables exist.
Since the mind is fed by innumerable streams and sources, it is difficult to pinpoint
how a particular good idea came to the researcher. The following are some of the
popularly known sources of research hypothesis:
Scientific theories: A systematic review and analysis of theories developed
in the field of psychology, sociology, economics, political science and
Self-Instructional
Material 51
Hypothesis biological science may provide the researcher with potential clues for
constructing a good and testable hypothesis.
Expert opinions: Discussion with the experts in the field of research may
further help the researcher obtain necessary insight and skill into the problem
NOTES
and in formulation of a hypothesis.
Method of related difference: When we find that two phenomena differ
constantly and the other circumstances remaining the same, we suspect a
causal connection. For example, when we find more uncontrolled traffic in
a locality, resulting in a greater number of road accidents, we suspect a
causal connection between uncontrolled traffic and road accidents. This
method also suggested a hypothesis.
Intellectual equipment of researcher: Intellectual abilities of a researcher
like creative thinking and problem solving techniques are very helpful in the
formulation of a good hypothesis.
Related literature: Related literature is the most important source of
hypothesis formulation. A review of this literature may reveal to the researcher
the variables that have been considered important in relation to his/her
problem, which aspects have already been studied and which still remain to
be studied, which theories have supported the relationships and which
theories present a contradictory relationship. Familiarity with related literature
may give the researcher a tremendous advantage in the construction of
hypothesis.
Experience: One’s own experience may be a rich source of hypothesis
generation. Personal experiences of an individual which has been gained
through reading of biographies, autobiographies, newspaper readings or
through informal talks among friends, etc., can be a potential source of
generation of a hypothesis. For example, a researcher who is working on
the effectiveness of guidance in teaching, can think of factors such as the
teacher’s polite behaviour, techniques of counselling, mastery over the
subject, effective use of teaching skills, decision-making capability, perception
of his/her competence, perception of student’s capacity for better interaction,
use of communication skills, etc.
Analogies: Several hypotheses in a branch of knowledge may be made by
using analogies from other sciences. Models and theories developed in a
discipline may help, through extrapolation, in the formulation of hypothesis
in another discipline. By comparing the two situations, analysing their
similarities and differences, some rationale may emerge in the mind of the
researcher which may take the form of a hypothesis for testing. For example,
in a research problem like the studying the factors of unrest among college
level students, the researcher insightfully thinks: ‘Why was unrest found
among school students? and What has changed them: quality of teaching or
quality of leadership?’
Self-Instructional
52 Material
Arguing analogically in this way may lead the investigator to some conclusions Hypothesis
which may be used for identifying variables and relationships, which form
the basis of hypothesis construction. If a researcher knows from previous
experience that the old situation is related to other factors Y and Z as well
as to X, he/she may reason out that the new situation may also be related to NOTES
Y and Z.
Methods of residues: When the greater part of a complex phenomenon
is explained by some causes already known, we try to explain the residual
part of phenomenon according to the known law of operation. It also
provides possible hypothesis.
Induction by simple enumeration: Sometimes scientists take common
experience as a starting point of their investigation. For example, after
observing a large number of scarlet flowers that are devoid of fragrance,
we frame a hypothesis that all scarlet flowers are devoid of fragrance. Thus
induction by simple enumeration is a source of discovery.
Formulation of hypothesis: It may also originate from the need and
practice of present times.
Existing empirical uniformities: In terms of common sense proposition,
the existing empirical uniformities may form the basis for scientific examination.
A study of general culture: It is also a good source of hypothesis.
Suggestions: When given by other researchers in their reports, suggestions
are quite helpful in establishment of hypothesis for future studies.
Self-Instructional
Material 53
Hypothesis Example: “There is a positive relationship between socio-economic status
and scholastic achievement.” This hypothesis stipulates that if socio-economic
status of students is high, scholastic achievement will also be high.
Self-Instructional
Material 55
Hypothesis The various points that emerge from the above discussion may thus be
summarised and restated as follows:
A statistical hypothesis, often called a null hypothesis H0, is tested against
an alternative hypothesis H1. The latter can be stated in different ways
NOTES
depending on the problem situation as to what the parameter is expected
to be at a given point of time.
A statistical hypothesis is stated with reference to a population parameter,
never in relation to the corresponding sample statistic. The appropriate
sample statistic serves merely as a means, by providing a point estimate,
to decide whether a statistical hypothesis is to be rejected or accepted.
A statistical hypothesis is always stated in the present tense in a manner
that some general state of affairs currently exists. Use of future tense in
stating a hypothesis is not admissible, as in that case the null hypothesis
will involve a future state of affairs about something that may not exist.
the basis for reporting the conclusions of the study on the basis of these
conclusions. The researcher can make the research report interesting and
meaningful to the reader. The importance of a hypothesis is generally
recognized more in the studies which aim to make predictions about some NOTES
outcome. In an experimental study, the researcher is interested in making
predictions about the expected outcomes and, hence the hypothesis takes
on a critical role. In the case of historical or descriptive studies, however,
the researcher investigates the history of an event, or life of a man, or seeks
facts in order to determine the status quo of a situation and hence may not
have a basis for making a prediction of the results. In studies of this nature,
where fact finding itself is the objective of the study, a hypothesis may not
be required.
Most historical or descriptive studies involve fact finding as well as the
interpretation of facts in order to draw generalizations. For all such major studies,
a hypothesis is recommended so as to explain observed facts, conditions or
behaviour and to serve as a guide in the research process. If a hypothesis is not
formulated, a researcher may waste time and energy in gathering extensive empirical
data, and then find that he/she cannot state facts clearly and detect relevant
relationships between variables as there is no hypothesis to guide him/her.
possible, research should proceed from a hypothesis. In the words of Van Dalen
(1973), ‘a hypothesis serves as a powerful beacon that lights the way for the
research worker’.
NOTES
3.7 HYPOTHESIS TESTING AND THEORY AND
ERRORS IN TESTING OF HYPOTHESIS
A claim or hypothesis about the values or population parameters is known as the
Null Hypothesis and is written as H0. In the case of the above discussed situation,
our assumption that a butler is innocent would form the null hypothesis and would be
stated as follows:
H0 = The butler is innocent
This hypothesis is then tested with the available evidence and the decision is
made whether to accept this hypothesis or reject it. If this hypothesis is rejected,
then we accept the alternate hypothesis which is that the butler is not innocent.
This alternate hypothesis is denoted as H1 and is stated as:
H1 = The butler is not innocent
The process involves testing of the null hypothesis. If the null hypothesis is
rejected, then the alternate hypothesis is accepted. It should be noted that the
acceptance of the alternate hypothesis does not mean that it is correct. It simply
means that there is not enough evidence to be reasonably sure that the null hypothesis
is acceptable.
As already explained, there are two types of errors that can be used in
making decisions regarding accepting or rejecting the null hypothesis. The first
type of error, known as Type I error is used when the null hypothesis is rejected
even if it is true. The second type of error, known as Type II error is used when a
null hypothesis is accepted even if it was not true and should have been rejected.
In statistical hypothesis testing and decision-making about the values of
population parameters as defined by the sample statistics, the null hypothesis asserts
that there is no true difference between the sample statistics and the corresponding
population parameter under consideration and if indeed there is any visible
difference, it is considered to be due to natural fluctuations in sampling.
To conclude we say that,
Null Hypothesis H0– An assertion about the population parameter that
is being tested by the sample results.
Alternate Hypothesis H1 – A claim about the population parameter
that is accepted when the null hypothesis is rejected.
Type I Error – An error made in rejecting the null hypothesis, when in fact
it is true.
Self-Instructional
Material 59
Hypothesis Type II Error – An error made in accepting the null hypothesis, when in
fact it is false.
Type I error is denoted by (Alpha) and is expressed as a probability of
rejecting a true hypothesis. It is also known as the level of significance. 1 –
NOTES
expresses the level of confidence. For example, = 0.05 means that the confidence
level is 95% or 0.95.
Type II error is denoted by (Beta) and is expressed as the probability of
accepting a false hypothesis. It is desirable to have the value as low as possible
for its value reflects the power of the test being performed and a low value
indicates that the test of significance is powerful and reliable.
3.7.1 Procedure for Hypothesis Testing
The general procedure for hypothesis testing consists of the following steps:
1. State the Null Hypothesis as well as the Alternate Hypothesis. This
means stating the assumed value of the population parameter which is to
be tested. For example, suppose that we want to test the hypothesis that
the average IQ of our college students is 130. Then this would become our
null hypothesis and the alternate hypothesis would be that this average IQ
is not 130. These statements are expressed as follows:
H0 : = 130
H1 : 130
2. Establish a Level of Significance Prior to Sampling. The level of
significance signifies the probability of committing Type I error and is
generally taken as equal to 0.05, which really means that after the hypothesis
has been tested and a decision is made, we will still be making an error
in rejecting the null hypothesis when in fact it is true, 5% of the time.
Sometimes the value is established as 0.01, but it is at the discretion of
the investigator to select its value, depending upon the sensitivity of the
study.
3. Determine a Suitable Test Statistic. This means the choice of appropriate
probability distribution to use with the particular available information under
consideration. The normal distribution using the Z score table or the
t-distribution is most often used.
4. Define the Rejection (Critical) Regions. The critical region will be
established on the basis of the choice of the value of the level of significance
. For example, if we select the value of = 0.05, and we use the
standard normal distribution as our test statistic for testing the population
parameter , then as we have discussed before, the difference between the
assumption of null hypothesis, assumed value of this population parameter
and the value obtained by the analysis of sample results is not expected to
Self-Instructional
60 Material
Hypothesis
be more than ± 1.96 X at = 0.05. This relationship can be shown in
Figure 3.1.
NOTES
In the above figure, if the sample X statistic falls within 1.96 X of the
assumed value of under the assumption of null hypothesis H0, then we
accept the null hypothesis as being correct at 95% confidence level (or
0.05 level of significance). The difference between X and which may be
any value between X1 and or X2 and is considered to be accidental
or due to chance element and is not considered significant enough or real
enough to reject null hypothesis, so that for all practical purposes the value
of X is considered equal to even though X can have any value between
X1 and X2 as shown above. However, if the value of X falls beyond X2
on the upper side or beyond X1 on the lower side, then this difference
between the values of X and would be considered significant and it will
lead to rejection of null hypothesis. Since 5% of the time, this difference
between the values of X and would be significant with 2.5% of the time
X being too far above (beyond X2) and 2.5% of the time being too far
below (below X1), the area of rejection will be on both sides of the
mean extending into the tail sections of the curve. This area of rejection is
known as the critical region.
5. Data Collection and Sample Analysis. This involves the actual collection
and computation of the sample data. A sample of the pre-established size
n is collected and the estimate of the population parameter is calculated.
This estimate is the value of the test statistic. For example, if we are testing
a hypothesis about the value of population mean , then the test statistic
would be the sample mean X . Then we test this statistic to check whether
it falls in the critical region or in the acceptance region. For example, if we
want to test for the average IQ of the college students to be 130, then in
that case we have to see that our population mean must be tested. We
take a random sample of a given size n and calculate its mean X and then
Self-Instructional
Material 61
Hypothesis
test it to see if the value of this X falls in the area of acceptance or in the
area of rejection at a given level of significance.
6. Making the Decision. Before the statistical decision is made, a decision
NOTES rule must be established. Such decision rule will form the basis on which
the null hypothesis will be accepted or rejected. This decision rule is really
a formal statement of the obvious purpose of the test. For example, this
rule could be stated as follows,
Accept the null hypothesis if the value of sample statistic X falls
within the area of acceptance, otherwise reject the null hypothesis.
Based upon this established decision rule, a decision can be made whether
to accept or reject the null hypothesis.
3.7.2 Committing Errors: Type I and Type II
Types of Errors: There are two types of errors in statistical hypothesis,
which are as follows:
o Type I Error: In this type of error, you may reject a null hypothesis
when it is true. It means rejection of a hypothesis, which should have
been accepted. It is denoted by (alpha), and is also known alpha
error.
o Type II Error: In this type of error, you are supposed to accept a null
hypothesis when it is not true. It means accepting a hypothesis, which
should have been rejected. It is denoted by (beta), and is also known
as beta error.
Type I error can be controlled by fixing it at a lower level, for example, If
you fix it at 2%, then the maximum probability to commit Type I error is 0.02. But
reducing Type I error, has a disadvantage when the sample size is fixed as it
increases the chances of Type II error. In other words, it can be said that both
types of errors cannot be reduced simultaneously. The only solution of this problem
is to set an appropriate level by considering the costs and penalties attached to
them or to strike a proper balance between both types of errors.
In a hypothesis test, a Type I error occurs when the null hypothesis is rejected
when it is in fact true; that is, H0 is wrongly rejected. For example, in a clinical trial
of a new drug, the null hypothesis might be that the new drug is no better, on
average, than the current drug; that is H0: there is no difference between the two
drugs on average. A Type I error would occur if we concluded that the two drugs
produced different effects when in fact there was no difference between them.
In a hypothesis test, a Type II error occurs when the null hypothesis H0, is
not rejected when it is in fact false. For example, in a clinical trial of a new drug,
the null hypothesis might be that the new drug is no better, on average, than the
current drug; that is H0: there is no difference between the two drugs on average.
A Type II error would occur if it were concluded that the two drugs produced the
Self-Instructional
62 Material
same effect, that is, there is no difference between the two drugs on average, Hypothesis
Accept H0 Reject H0
The level of significance implies the probability of Type I error. A five per cent level
implies that the probability of committing a Type I error is 0.05. A one per cent level
implies 0.01 probability of committing Type I error.
Lowering the significance level and hence the probability of Type I error is
good but unfortunately it would lead to the undesirable situation of committing
Type II error.
To sum up:
Type I Error: Rejecting H0 when H0 is true.
Type II Error: Accepting H0 when H0 is false.
Note: The probability of making a Type I error is the level of significance of a statistical test.
It is denoted by .
The probability of making a Type II error is denoted by .
3.9 SUMMARY
Self-Instructional
64 Material
A statistical hypothesis stated with a view to testing its validity is in fact a null Hypothesis
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Self-Instructional
66 Material
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils, Hypothesis
Self-Instructional
Material 67
Sampling Techniques
4.0 INTRODUCTION
The main aim of research is to discover principles that have universal application.
Generally, research in education includes all such assumptions that are based on a
large number of samples/units/objects. It would be impractical if not impossible to
test or observe each unit of population under controlled conditions in order to
arrive at principles having universal validity. A ‘population’ is any group of
individuals/units that have one or more characteristics in common which are of
interest to the researcher, for a particular research. A ‘sample’ is a small percentage
of the larger group who are selected for research. A sample can be statistically
explained as being a subset of a population. The sample will be able to give an
idea of the characteristics of the larger group from where it has been drawn. It is
possible to make deductions about the larger population on the basis of the sample.
For selecting a sample, it is necessary to have a sampling frame. After defining
a population and listing all the units, a researcher selects a sample of units from the
sampling frame. Sampling design refers to a definite plan for obtaining a sample
from the sampling frame. It refers to the technique or procedure, which a researcher
adopts in selecting some sampling units from where inferences about population
are drawn. An error in statistics is the difference between the value of a statistic
and that of the corresponding parameter. These errors arise due to chance differences
between the members of population included in the sample and those not included.
This unit discusses the concept of population and sample, methods of sampling,
sampling design, sampling distribution and sampling errors.
Self-Instructional
68 Material
Sampling Techniques
4.1 OBJECTIVES
Self-Instructional
70 Material
Where it is necessary to have the exact and accurate result and a slight Sampling Techniques
Self-Instructional
72 Material
As the scale of operation involved in a sample survey is small, the quality of Sampling Techniques
interviewing, supervision and other related activities is better than the census
survey.
Sampling provides adequate information needed for the purpose and is NOTES
sufficiently reliable for surveys.
Samples can be of different types. The following are the characteristics of a good
sample:
True representative: A good sample is a true representative of the
population corresponding to its properties.
Free from bias: A good sample does not permit prejudices, pre-conceptions
and imagination to influence its choice.
Comprehensive: Comprehensiveness is a quality of a sample which is
controlled by the specific purpose of the investigation. A sample may
be have all the traits required, but still not be a good representative of
population.
Economical: A sample should be economical from energy, time and money
viewpoint.
Approachable: A sample should be easily approachable. The research
tools can be easily administered on them.
Good size: Size of a sample should be such that it yields an accurate result.
The probability of error can be estimated.
Feasible: A good sample makes the research more feasible.
Practical: A good sample has the practicability for research situations.
Objective: This refers to objectivity in selecting a sampling procedure or
absence of subjective elements from situation.
Accurate: A good sample yields accurate estimates of statistics and does
not allow for errors.
Self-Instructional
Material 73
Sampling Techniques
4.6 TECHNIQUES OF SAMPLING (PROBABILITY
AND NON-PROBABILITY SAMPLING
TECHNIQUES)
NOTES
The sampling method was used in social sciences research in early 1754 by
A.L. Bowley. Since then the method has been progressively used. The sampling
methods are broadly classified into two types: (i) Probability Sampling and
(ii) Non-Probability Sampling.
Criteria for Selecting Sampling
The type of sample has to be selected depending on the area of research. For this
purpose, a variety of sampling methods can be employed, individually or in
combination. Factors commonly influencing the choice of the sampling design
include:
Nature and quality of the frame.
Availability of supporting information about units on the frame.
Accuracy requirements and the need to measure accuracy.
Whether detailed analysis of the sample is expected.
Cost/operational concerns.
As there are various sampling methods, it becomes crucial to select an appropriate
sampling method. Young has suggested the following three criteria to be considered
while selecting a sampling method:
A measurable or known probability sampling technique should be used to
control the risk of errors in the sample estimate.
Simple, straightforward and workable methods adapted to available facilities
and personnel should be used.
Achieving optimum balance between expenditure incurred and maximum
of reliable information should be the guiding principle.
The decision whether a probability sampling or a non-probability sampling is to be
applied rests on the constraints which are not very different from those stated
earlier. These are: (i) objectives of the study, (ii) type of study, and (iii) availability
of the resources for the study.
(i) If the objective of the research is to apply the results of the study to a small
local group then sampling may not be given as much consideration as in a
study where the results are to be applied to a larger group. Action research
generally does not require sampling from a larger group. Most of the time
sampling is not very essential in historical research, whereas survey studies
generally have a more rigorous sampling.
Self-Instructional
74 Material
(ii) The availability of time, funds, manpower and equipment required is another Sampling Techniques
Self-Instructional
Material 75
Sampling Techniques are equally applicable across racial groups, SRS cannot be used as random selection
of a sub-group as it may not provide accurate findings.
Theoretically, this is a method of selecting ‘n’ units from N units in such a
way that everyone in the population of N units has an equal chance of being selected.
NOTES
This can be done through the following steps:
(i) Defining population by specifying its various limits.
(ii) Preparing the sampling frame.
(iii) Incorporating the names or serial numbers of individual units in the
sampling frame (every unit is to be listed, order does not make any
difference).
It is important to re-emphasize here that a random sample is not necessarily
an identical representation of the population. After this, to get the required ‘n’
units different techniques are available. These techniques are discussed below.
(a) Lottery method: After numbering every unit in the population, they
are well mixed. The required numbers of units are then drawn from all
these well mixed units. The individuals’ objects with these identification
named numbers are then picked up for inclusion in the sample.
However, this technique has some objections. When the population is
very large and includes such individuals/objects which are of such
nature that could not be mixed and further if ‘well mixing’ is not attained
despite all efforts, the principle of randomness in the population may
be violated.
(b) Random table method: The use of random numbers or manual lot
drawing will be too cumbersome to recommend in case of a large
population. In such situations, computer generated random selection
should be resorted, in order to save time and labour. Tables of random
numbers have been generated by computers producing a random
sequence of digits, e.g., random digit tables by Rand Corporation
and prepared by Kendall & Smith, by Fisher & Yates and by Tippet
are frequently used. The required numbers of units are selected from
such a table in any convenient and systematic way. Now suppose we
have to select 20 distance learners for interviews from 80 distance
learners registered at a study centre. We may start with any column
and any row. Because we want 20 numbers, i.e., two digit numbers,
we have to select only the first two digits from each number. If we
select the first column and start from first row then we will get following
22 digit numbers—23, 05, 14, 38, 97, 11, 43, ...............,. 61. You
will notice that numbers greater than 80 will have to be deleted from
this list and for the remaining numbers selecting any other column and
the row the procedure will have to be repeated, till we get the required
number, i.e., 20. If any number is repeated in this list, it is to be
Self-Instructional
76 Material
substituted by selecting the next number. Until a sample of desired Sampling Techniques
Self-Instructional
78 Material
Disadvantages of systematic sampling Sampling Techniques
Self-Instructional
82 Material
Advantages of cluster sampling Sampling Techniques
Self-Instructional
Material 83
Sampling Techniques used for it. A teacher-educator, e.g., may select the students from a school situated
in the same campus which serves as a practising school for the concerned college
of education, find the effectiveness of concept attainment model to teach a
mathematical concept say, a quadrilateral.
NOTES
Advantages of incidental sampling
The administrative convenience of obtaining samples for the study, the ease of
testing, saving in time and completeness of the data collected are some of the
merits of this method.
Disadvantages of incidental sampling
Since there is no well-defined population and no random sampling method is applied
to select the sample, the standard error formulae are applied with a high degree of
approximation. Hence, no valid generalization can be drawn. Any attempt at
generalization based on such data and conclusion thereof will be misleading.
Judgment sampling
It involves selection of groups from the population on the basis of available
information. The groups should be representative of the population. It has good
evidence and is based on experience. It is an economical method.
Purposive sampling
Another non-probability sampling method is ‘purposive sampling’. The sample is
selected by some arbitrary method because it is known to be representative of the
total population, or it is known that it will produce well matched groups. The idea
is to pick out the sample in relation to some criterion which is considered important
for the particular study.
In this method, samples are chosen because they resemble some larger
group with respect to one or more characteristics. The controls of criteria for
categorization in such samples are usually identified as representative areas, such
as a state, a district, a city, etc., or representative characteristics of individuals,
such as age, sex, socio-economic status, etc., or representative types of groups,
such as elementary school teachers, secondary school teachers, college teachers,
university teachers, etc. These controls criteria may be further sub-divided, e.g.,
the group of college teachers can be divided into male and female teachers or
teachers in science/arts/commerce colleges, etc.
It has to be noticed here that up to this stage the controls are somewhat
similar to stratification criteria. After deciding upon the category required for the
research, the researcher has to select the sample. Actual selection of the units for
inclusion in the sample is done purposively and not randomly, e.g., in order to
tackle the problem of indiscipline only the undisciplined students are selected as
the sample, excluding others on the basis of past experience.
Self-Instructional
84 Material
Advantages of purposive sampling Sampling Techniques
Sampling Errors
Even if utmost care has been taken in selecting a sample, the results derived from
a sample study may not be exactly equal to the true value in the population. The
Self-Instructional
Material 85
Sampling Techniques reason is that estimate is based on a part and not on the whole and samples are
seldom, if ever, perfect miniature of the population. Hence, sampling gives rise to
certain errors known as ‘sampling errors’ or sampling fluctuations.
In other words, a sample survey requires study in small portions of population
NOTES
as there can be certain amount of inaccuracy in the information collected during
sampling analysis. This inaccuracy is called sampling error or error variance.
Sampling errors are those errors, which arise on account of sampling and generally
happen to be random variations in the sample estimates of the actual population
values. Figure 4.1 shows sampling error.
Sampling errors occur randomly and are equally likely to be in either direction
and the magnitude of sampling error depends on the nature of the universe. The
more uniform the universe is, the smaller is the sampling error. Sampling error is
inversely proportional to the size of the sample and vice-versa. In addition, sampling
error is the product of the critical value at a certain level of significance and the
standard error.
Sampling Error = Frame Error + Chance Error + Response Error
Sampling errors would not be present in a complete enumeration survey.
However, the errors can be controlled. The modern sampling theory helps in
designing the survey in such a manner that the sampling errors can be made
insignificant. Sampling errors are of two types: (i) biased and (ii) unbiased.
These errors arise from any bias in selection, estimation, etc. For example,
if in place of simple random sampling, if deliberate sampling has been used in a
particular case; some bias is introduced in the result, and hence such errors are
called ‘biased sampling errors’.
Self-Instructional
86 Material
An error in statistics is the difference between the value of a statistic and Sampling Techniques
that of the corresponding parameter. These errors arise due to chance differences
between the members of population included in the sample and those not included.
Thus, the total sampling error is made up of errors due to bias, if any, and
NOTES
the random sampling error. The essence of bias is that it forms a constant component
of error that does not decrease in a large population as the number in the sample
increases. Such error is, therefore, also known as ‘cumulative/non-compensating
error’. The random sampling error, on the other hand, decrease as an average as
the size of the sample increases. Such error is, therefore, also known as ‘non-
cumulative/compensating error’.
Bias may arise due to: (i) faulty process of selection, (ii) faulty work during
the collection, and (iii) faulty methods of analysis.
Faulty selection of the sample may give rise to bias in a number of ways.
Some of which are discussed below:
(a) Deliberate selection The deliberate selection of a ‘representative’ sample.
(b) Conscious/Unconscious bias in the selection of ‘random’ sample: The
randomness of selection may not really exist, even though the investigator
claims that he/she had a random sample if he/she allows his/her desire to
obtain a certain result to influence his/her selection.
(c) Substitution: Substitution of an item in place of one chosen in random
sample sometimes leads to bias. Thus, if it were decided to interview every
50th household in a street, it would be inappropriate to interview the 51st
or any other number in its place as the characteristics possessed by it will
differ from those which were originally to be included in the sample.
(d) Non-response: If all the items to be included in the sample are not covered
then, there will be bias even though no substitution has been attempted.
This fault particularly occurs in mailed questionnaires, which are incompletely
returned. Moreover, the information supplied by the informants may also
be biased.
(e) An appeal to the vanity: An appeal to the vanity of the person questioned
may give rise to yet another kind of bias. For example, the question ‘Are
you a good student?’ is such that most of the students would answer ‘yes’.
Any consistent error in measurement will give rise to bias whether the
measurements are carried out on a sample or on all the units of the population.
The danger of error is, however, likely to be greater in sampling work, since the
units measured are usually smaller.
Bias may arise due to improper formulation of the decision problem or
wrongly defining the population, specifying the wrong decision, securing an
inadequate frame, and so on. Biased observations may result from a poorly designed
questionnaire, an ill-trained interviewer, failure of a respondent’s memory, etc.
Self-Instructional
Material 87
Sampling Techniques Bias in the flow of data may be due to unorganized collection procedure, faulty
editing or coding of responses.
In addition to bias which arises from faulty process of selection and faulty
collection of information, faulty methods of analysis may also introduce bias. Such
NOTES
bias can be avoided by adopting the proper methods of analysis.
If possibilities of bias exist, fully objective conclusions cannot be drawn.
The first essential of any sampling or census procedure must, therefore, be the
elimination of all sources of bias. The simplest and the only certain way of avoiding
bias in the selection process is for the sample to be drawn either entirely at random
or subject to restrictions, which while improving the accuracy are of such a nature
that they do not introduce bias in the results. In certain cases, systematic selection
may also be permissible.
Once the absence of bias has been ensured, attention should be given to the
random sampling errors. Such errors must be reduced to the minimum so as to
attain the desired accuracy.
Apart from reducing errors of bias, the simplest way of increasing the
accuracy of a sample is to increase its size. The sampling error usually decreases
with increase in sample size and in fact in many situations the decrease is inversely
proportional to the square root of the sample size. Figure 4.2 illustrates the increase
and decrease proportion between sampling error ad sample size.
From Figure 4.2, it is clear that though the reduction in sampling error is
substantial for initial increases in sample size, it becomes marginal after a certain
stage. In other words, considerably great effort is needed after a certain stage to
decrease the sampling error than in the initial instances. Hence after that stage
sizable reduction in cost can be achieved by lowering even slightly the precision
required.
From this point of view, there is a strong case for resorting to a sample
survey to provide estimates within permissible margins of error instead of a complete
Self-Instructional
88 Material
enumeration survey, as in the latter the effort and the cost needed will be substantially Sampling Techniques
Self-Instructional
Material 89
Sampling Techniques These sources are not exhaustive, but are given to indicate some of the
possible sources of error. In a sample survey, non-sampling errors may also arise
due to defective frame and faulty selection of sampling units.
In some situations, the non-sampling errors may be large and deserve greater
NOTES
attention than the sampling errors. While, in general sampling errors decrease with
increase in sample size, non-sampling errors tend to increase with the sample size.
In the case of complete enumeration, non-sampling errors and in the case of
sample surveys, both sampling and non-sampling errors require to be controlled
and reduced to a level at which their presence does not vitiate the use of final
results.
The reliability of samples can be tested in the following ways:
(i) More samples of the same size should be taken from the same universe and
their results be compared. If results are similar, the sample will be reliable.
(ii) If the measurements of the universe are known then they should be compared
with the measurements of the sample. In case of similarity of measurement,
the sample will be reliable.
(iii) Sub-samples should be taken from the samples and studied. If the results of
sample and sub-sample study show similarity, the sample should be
considered reliable.
1. A ‘sample’ is a small percentage of the larger group who are selected for
research. A sample can be statistically explained as being a subset of a
population. The sample will be able to give an idea of the characteristics of
the larger group from where it has been drawn. It is possible to make
deductions about the larger population on the basis of the sample.
2. Sampling is the process of obtaining information about an entire population
by examining only a part of it. Sampling is required as it saves time and
money, it produces results at a faster speed, it enables more accurate
measurement for a sample study as it is conducted by experienced
investigators and it is the only method for an infinitely large population.
Self-Instructional
90 Material
3. Sampling methods are broadly classified into two types: (i) Probability Sampling Techniques
Self-Instructional
Material 91
Sampling Techniques
4.9 SUMMARY
Self-Instructional
92 Material
Systematic sampling is a variant of the random process of sampling. In this Sampling Techniques
technique of the requisite number of sample units are selected from the
population.
To increase the precision, ‘stratified sampling’ can be one option. The term
NOTES
‘stratified’ is very much self-explanatory. It involves dividing the population
into such sub-populations (strata) that each one of them is homogeneous
within itself. Population can be divided into different categories by using
these ‘strata’ or layers. Each stratum or section is then treated independently
as a separate sub-group.
Double sampling is a type of sampling which includes both questionnaire
and interview methods for probing a research problem.
The main distinction between the multi-stage and the multi-phase sampling
is the use of unit of sampling at different levels in multi-stage sampling but
not in multi-phase sampling.
In Cluster sampling, the units of samples close to each other are chosen in
clusters, for example, households in the same street or successive items of
a production-line. The population is divided into clusters and some of them
are chosen randomly. Then, the clustered units are selected using random
sampling method.
The non-probability sampling methods are based on the judgments of the
investigator as the most important elements of control. The guiding principles
in non-probability methods are— availability of the subjects, the personal
judgment of the investigator, and convenience in carrying out the research.
The non-probability sampling methods are of following types, namely
incidental or accidental sampling, judgement sampling, purposive sampling,
quota sampling and snowball sampling.
Even if utmost care has been taken in selecting a sample, the results derived
from a sample study may not be exactly equal to the true value in the
population. The reason is that estimate is based on a part and not on the
whole and samples are seldom, if ever, perfect miniature of the population.
Hence, sampling gives rise to certain errors known as ‘sampling errors’ or
sampling fluctuations.
In other words, a sample survey requires study in small portions of population
as there can be certain amount of inaccuracy in the information collected
during sampling analysis. This inaccuracy is called sampling error or error
variance.
The sampling error usually decreases with increase in sample size and in
fact in many situations the decrease is inversely proportional to the square
root of the sample size.
Self-Instructional
Material 93
Sampling Techniques The behaviour of the non-sampling errors with increase in sample size is
likely to be opposite of that of sampling error, that is, the non-sampling
error is likely to increase with increase in sample size.
Non-sampling errors can occur at every stage of planning and execution of
NOTES
the census or survey. Such errors can arise due to a number of causes, such
as defective methods of data collection and tabulation, faulty definition,
incomplete coverage of the population or sample, etc.
In some situations, the non-sampling errors may be large and deserve greater
attention than the sampling errors. While, in general sampling errors decrease
with increase in sample size, non-sampling errors tend to increase with the
sample size.
Self-Instructional
94 Material
Long Answer Questions Sampling Techniques
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
Material 95
Research Tools
BLOCK - II
RESEARCH TOOLS AND DIFFERENT TYPES
OF RESEARCH
NOTES
5.0 INTRODUCTION
5.1 OBJECTIVES
collection
Describe the processing of the techniques used in data collection
Analyse the problems related to the classification of data NOTES
Know the meaning and usage of rating scale and attitude scale
Understand the method of writing a research proposal
Let us study the various tools and techniques used for data collection.
5.2.1 Observation
Observations have lead to some of the most important scientific discoveries in
human history. Charles Darwin used his observations of animal and marine life at
the Galapagos Islands to help him formulate his theory of evolution that he described
in On the Origin of Species. Today, social scientists, natural scientists, engineers,
computer scientists, educational researchers and many others use observations as
a primary research method.
The kind of observations one makes depends on the subject being
researched. Traffic or parking patterns on a campus can be observed to ascertain
what improvements could be made. Clouds, plants or other natural phenomena
can be observed as can people, though in the case of the latter one may often have
to ask for permission so as to not violate any privacy issue.
Observation may be defined as ‘a process in which one or more persons
monitor some real-life situation and record pertinent occurrences’. It is used
to evaluate the overt behaviour of the individual in controlled and uncontrolled
situations.
According to Jahoda: ‘Observation method is a scientific technique to
the extent that it (a) serves a formulated research purpose, (b) is planned
systematically rather than occurring haphazardly, (c) is systematically
recorded and related to more general propositions than presented as a set of
interesting curious, and (d) is subjected to checks and controls with respect
to validity, reliability, and precision much as is all other scientific evidence.’
According to Good and Hatt: ‘Observation may take many forms and is
at once the most primitive and the most modern of research techniques. It
includes the most casual, uncontrolled experiences as well as the most exact
film records of laboratory experimentation.’
Self-Instructional
Material 97
Research Tools Types of Observation
Observations are mainly classified in the following two types:
Participant observation: In the process of ‘participant observation’, the
NOTES observer becomes more or less one of the group members and may actually
participate in some activity or the other of the group. The observer may
play any one of the several roles in observation, with varying degrees of
participation, as a visitor, an attentive listener, an eager learner or as a
participant observer.
Non-participant observation: In the process of ‘non-participant
observation’, the observer takes a position where his/her presence is not
felt by the group. He/She may follow the behaviour of an individual or
characteristics of one or more groups closely. In this type of observation, a
one-way ‘vision screen’ permits the observer to see the subject but prevents
the subject from seeing the observer.
Observations may also be classified into the following categories:
i. Natural observation: Natural observation involves observing the
behaviour in a normal setting and in this type of observation; no efforts
are made to bring any type of change in the behaviour of the observed.
Improvement in the collection of information can be done with the
help of natural observations.
ii. Subjective and objective observation: All observations consist of
two main components, the subject and the object. The subject refers
to the observer, whereas the object refers to the activity or any type
of operation that is being observed. Subjective observation involves
the observation of one’s own immediate experience, whereas the
observations involving an observer as an entity apart from the thing
being observed are referred to as the ‘objective observation’.
Objective observation is also known as the ‘retrospection’.
iii. Direct and indirect observation: With the help of the direct method
of observation, one comes to know how the observer is physically
present, in which type of situation is he/she present and then this type
of observation monitors what takes place. Indirect method of
observation involves studies of mechanical recording or the recording
by some of the other means like photographic or electronic. Direct
observation is relatively straightforward as compared to indirect
observation.
iv. Structured and unstructured observation: Structured observation
works according to a plan and involves specific information of the
units that are to be observed and also about the information that is to
be recorded. The operations that are to be observed and the various
features that are to be noted or recorded are decided well in advance.
Self-Instructional Such observations involve the use of special instruments for the purpose
98 Material
of data collection that are also structured in nature. But in the case of Research Tools
Self-Instructional
100 Material
It should be recorded immediately. Research Tools
Self-Instructional
Material 101
Research Tools Planning effective observation
This includes the following:
Sampling to be observed should be adequate. There should be an appropriate
NOTES group of subjects.
Units of behaviour should be defined as accurately as possible.
Method of recording should be simplified.
Detailed instructions may be given to observers to eliminate the difference
in perspective of observers.
Too many variables may not be observed simultaneously.
Excessively long periods of observation without interspersed rest periods
should be avoided.
Observers should be fully trained.
Observers should be well equipped.
Conditions of observation should remain constant.
Number of observations should be adequate.
Records of observation must be comprehensive.
Length of each observation period, interval between periods, and number
of periods should be clearly stated.
Interpretations should be carefully made.
Disadvantages of Observation
The disadvantages of observation are as follows:
It is very difficult to establish the validity of observations.
Many items of observation cannot be defined.
The problem of subjectivity is involved.
Observation may give undue stress to aspects of limited significance simply
because they can be recorded easily, accurately and objectively.
Various observers observing the same event may concentrate on different
aspects of a situation.
The observer has little control over the physical situation.
Children being observed become conscious and begin to behave in an
unnatural manner.
Many children try to pose and exhibit at the time of observation.
There are certain situations which the observer is not allowed to observe,
and he is helpless in that way to produce an accurate account.
Self-Instructional
102 Material
It may not be feasible to classify all the events to be observed. Research Tools
NOTES Clinical interview: Such an interview follows after the diagnostic interview.
It is a means of introducing the patient to therapy.
Research interview: Research interview is aimed at getting information
required by the investigator to test his/her hypothesis or solve his/her problems
of historical, experimental, survey or clinical type.
Single interview or panel interviews: For the purpose of research, a
single interviewer is usually present. In case of selection and treatment
purposes, panel interviews are held.
Directed interview: It is structured, includes questions of the closed type
and is conducted in a prepared manner.
Non-directive interview: It includes questions of the open-end form and
allows much freedom to the interviewee to talk freely about the problem
under-study.
Focused interview: It aims at finding out the responses of individuals to
exact events or experiences rather than on general lines of enquiry.
Depth interview: It is an intensive and searching type of interview. It
emphasizes certain psychological and social factors relating to attitudes,
emotions or convictions.
It may be observed that on occasions several types are used to obtain the
needed information.
Other classifications of interviews are as follows:
Intake interview, as the initial stage in clinic and guidance centres.
Brief talk contacts as in schools and recreation centres.
Single hour interview.
Clinical psychological interview, stressing psychotherapeutic counselling
and utilizing case history data and active participation by the counsellor
in the re-education of the client.
Psychiatric interviews, similar to psychological counselling, but varying
with the personality and philosophical orientation of the individual worker
and with the setting in which used.
Psychoanalytic interviews.
The interview form of test.
Group interviews for selecting applicants for special course.
Research interview.
Self-Instructional
104 Material
Important Elements of Research Interview Research Tools
Self-Instructional
Material 105
Research Tools iv. Reporting the response
There are two chief means of recording opinion during the interview. If the question
is preceded, the interviewer need only check a box or circle or code, or otherwise
NOTES indicate which code comes closest to the respondent’s opinion. If the question is
not preceded, the interviewer is expected to record the response verbatim.
The following points may be kept in view in this respect:
Quote the respondents directly, just as if the interviewers were newspaper
reporters taking down the statement of an important official without
paraphrasing the reply, summarizing it in the interviewer’s own words,
‘polishing up’ any slang or correcting bad grammar that distorts the
respondent’s meaning and emphasis.
Ask the respondent to wait until the interviewer gets down ‘that last
thought’.
Do not write as soon as you have asked the question and do not write
while the respondent talks. Wait until the response is completed.
Use common abbreviations.
Do not record and evaluate the responses simultaneously.
v. Closing the interview
It should be accompanied by an expression of thanks in recognition of the
respondent’s generosity in sparing time and effort.
vi. Use of tape recorder in interview
It reduces the tendency of the interviewer to make an unconscious selection
of data favouring his/her biases.
The tape recorded data can be played more than once, and thus it permits
a thorough study of the data.
Tape recorder speeds up the interview process.
Tape recorder permits the recording of some gestures.
The tape recorder permits the interviewer to devote full attention to the
respondent.
No verbal productions are lost in a tape recorded interview.
Other things being equal, the interviewer who uses a tape recorder is able
to obtain more interviews during a given time period than an interviewer
who takes notes or attempts to reconstruct the interview from memory
after the interview has been completed.
Self-Instructional
106 Material
Indifferent Attitude of the Respondent and the Role of the Research Research Tools
Worker
It is observed that the research worker is likely to encounter several problems
arising out of the apathy of the respondents. In such a situation, the following NOTES
points may be kept in view:
i. When the respondent is really busy and has no time, the field worker may
request for a more convenient time.
ii. When the respondent simply wants to avoid the interview and is not inclined
to be bothered about it, the field worker should try to explain to him/her the
importance of the study, and how his/her own response is of material value
in the case.
iii. When the respondent is afraid to give the interview as it affects his/her boss
or the party to which he/she belongs or any other cause which is likely to
harm his/her interest, the field worker must assure the respondent that
absolute secrecy would be maintained by the researcher and the organization.
iv. When the respondent does not hold a high opinion about the outcome of
such interviews in general, or has a poor opinion about the research
organization or institution conducting it, it is the duty of the research worker
at such times to explain to him/her the importance of the problem, and
convince him/her regarding the status of the research body.
v. When the respondent is suspicious and he/she thinks that the enquiry is
either from the income tax department or some other secret agency, at such
times he/she may generally ask such questions. Who are you? Who told
you our name? Have you interviewed the neighbour?, etc. The research
worker should try to eliminate his/her suspicion. A letter of authority, the
letter head or the seal of the research body would prove to be useful on
such occasions.
vi. When the respondent is unsocial or otherwise confined to his/her own family
(such a tendency is mostly found in the case of newly married couples), the
research worker at such times will try to create his/her interest in the subject
of investigation.
vii. When the respondent is too haughty and thinks it below his/her dignity to
grant an interview to petty research workers, the investigator should get a
letter of introduction from an influential person.
Advantages of Interview Method
The advantages of interview method over other techniques are as follows:
A well-trained interviewer can obtain more data and greater clarity by altering
the interview situation. This cannot be done in a questionnaire.
Self-Instructional
Material 107
Research Tools An interview permits the research worker to follow-up leads as contrasted
with the questionnaire.
Questionnaires are often shallow and they fail to dig deeply enough to provide
a true picture of opinions and feelings. The interview situation usually permits
NOTES
much greater depth.
It is possible for a skilled interviewer to obtain significant information through
motivating the subject and maintaining rapport, other methods do not permit
such a situation.
The respondents when interviewed may reveal information of a confidential
nature which they would not like to record in questionnaire.
Interview techniques can be used in the case of children and illiterate persons
who cannot express themselves in writing. This is not possible in a
questionnaire.
The percentage of response is much higher than in case of a mailed
questionnaire.
Removal of misunderstanding: The field worker is personally present to
remove any doubt or suspicion regarding the nature of enquiry or meaning
of any question or term used. The answers are, therefore, not biased because
of any misunderstanding.
Creating a friendly atmosphere: The field worker may create a friendly
atmosphere for proper response. He/She may start a discussion, and develop
the interest of the respondent before showing the schedule. A right
atmosphere is very conducive for getting correct replies.
Possible to secure confidential interview: The interviewee may disclose
personal and confidential information which he/she would not ordinarily
place in writing on paper. The interviewee may need the stimulation of
personal contacts in order to be drawn out.
Advantages of clues: The interview enables the investigator to follow-up
leads and to take advantage of small clues, in dealing with complex topics
and questions.
Permits exchange of ideas: The interview permits an exchange of ideas and
information. It permits ‘give and take’.
Useful in the case of some categories of persons: The interview enables the
interviewee to deal with young children, illiterates and those with limited
intelligence or in who’s state of mind is not quite normal.
Useful apart from research purposes: Interviews are also used for pupil
counselling, for selection of candidates for instructional purposes, for
employment, for psychiatric work, etc.
Possibility of asking supplementary questions: The respondent does not
feel tired or bored. Supplementary questions may be put to enliven the
whole discussion.
Self-Instructional
108 Material
Avoiding handwriting: The difficulties of bad handwriting of the respondent, Research Tools
use of pencil, etc., are also avoided as every schedule is filled in by the
interviewer.
A probe into life pattern is possible: The personal contact with the respondent
NOTES
enables the field worker to probe more deeply into the character, living
conditions and general life pattern of the respondent. These factors have a
great bearing in understanding the background of any reply.
Reliable information: The information gathered through interviews has been
found to be fairly reliable.
Deeper probe: It is possible for the interviewer to probe into attitudes,
discover the origin of the problem, etc.
Interview technique is very close to the teacher: It is generally accepted that
no research technique is as close to the teacher’s work as the interview.
Possibility of repetition: Sometimes interviews can be held at suitable intervals
to trace the development of behaviour and attitudes.
Useful for several purposes: Interviews can be used for student counselling,
occupational adjustment, selection of candidates for educational courses,
etc.
Wide applicability: Interviews can be used for all kinds of research
methods—normative, historical, experimental, case studies and clinical
studies.
Cross questioning: Interview techniques provide scope for cross questioning.
Command of the interviewer: This technique allows the interviewer to remain
in command of the situation throughout the investigation.
Wider opportunities to know the interviewee: Through the respondent’s
incidental comments, facial expression, bodily movements, gestures, etc.,
an interviewer can acquire information that could not be obtained easily by
other means.
Useful for judging frankness, etc.: Cross questioning by the interviewer can
enable him/her to judge the sincerity, frankness and insight of the interviewee.
Disadvantages of Interview Method
The method of interview, in spite of its numerous advantages has the following
limitations:
Very costly: It is a very costly affair. The cost per case is much higher in this
method than in case of mailed questionnaires. Generally speaking, the cost
per questionnaire is much less than the cost per interview. A large number of
field workers may have to be engaged and trained in the work of collection
of data. All this entails a lot of expenditure and a research worker with
limited financial means finds it very difficult to adopt this method.
Self-Instructional
Material 109
Research Tools Biased information: The presence of the field worker while encouraging the
respondent to reply, may also introduce a source of bias in the interview. At
times the opinion of the respondent is influenced by the field worker and his
replies may not be based on what he thinks to be correct but what he thinks
NOTES the investigator wants.
Time consuming: It is a time consuming technique as there is no guarantee
how much time each interview can take, since the questions have to be
explained, interviewees have to assured and the information extracted.
Expertness required: It requires a high level of expertise to extract information
from the interviewee who may be hesitant to part with this knowledge.
Among the important qualities to be possessed by an interviewer are
objectivity, insight and sensitivity.
5.2.3 Questionnaire
Questionnaire Tools
A questionnaire is ‘a tool for research, comprising a list of questions whose answers
provide information about the target group, individual or event’. Although they are
often designed for statistical analysis of the responses, this is not always the case.
This method was the invention of Sir Francis Galton. Questionnaire is used when
factual information is desired. When opinion rather than facts are desired, an
opinionative or attitude scale is used. Of course, these two purposes can be
combined into one form that is usually referred to as ‘questionnaire’.
Questionnaire may be regarded as a form of interview on paper. The
procedure for the construction of a questionnaire follows a pattern similar to that
of the interview schedule. However, because the questionnaire is impersonal, it is
all the more important to take care of its construction.
A questionnaire is a list of questions arranged in a specific way or randomly,
generally in print or typed and having spaces for recording answers to the questions.
It is a form which is prepared and distributed for the purpose of securing responses.
Thus a questionnaire relies heavily on the validity of the verbal reports.
According to Goode and Hatt, ‘in general, the word questionnaire refers
to a device for securing answers to questions by using a form which the
respondent fills himself’.
Barr, Davis and Johnson define questionnaire as, ‘questionnaire is a
systematic compilation of questions that are submitted to a sampling of
population from which information is desired’ and Lundberg says,
‘fundamentally, questionnaire is a set of stimuli to which literate people are
exposed in order to observe their verbal behaviour under these stimuli’.
Self-Instructional
110 Material
Types of Questionnaire Research Tools
Figure 5.1 depicts the types of questionnaires that are used by researchers.
Questionnaire NOTES
Closed form Open form Pictorial
Individual
In groups
In print
Projected
investigator can establish a rapport with the respondents; (ii) the purpose of
the questionnaire can be explained; (iii) the meaning of the difficult terms
and items can be explained to the respondents; (iv) group administration
when the respondents are available at one place is more economical in time NOTES
and expense; (v) the proportion of non-response is cut down to almost
zero; and (vi) the proportion of usable responses becomes larger. However,
it is more difficult to obtain respondents in groups and may involve
administrative permission which may not be forthcoming.
Computerized questionnaire: It is the one where the questions need to
be answered on the computer.
Adaptive computerized questionnaire: It is the one presented on the
computer where the next questions are adjusted automatically according to
the responses given as the computer is able to gauge the respondent’s ability
or traits.
Appropriateness of Questionnaire
The qualities and features which make questionnaires an effective instrument of
research and help to elicit maximum information are discussed below:
Type of information required: The usefulness and effectiveness of a
questionnaire is determined by the kind of information sought. Not every
type of questionnaire can be elicited through it. A questionnaire which will
consume more than 10–20 minutes is unlikely to be responded to well.
Also, the questions should be explicit and capable of clear-cut replies.
Type of respondent reached: A good deal depends upon the types of
respondents covered by the questionnaire. All types of individuals cannot
be good respondents. Only literate and socially conscious individuals would
give any consideration to a questionnaire. Also, the respondent must be
competent to answer the kind of questions contained in a particular
questionnaire.
Accessibility of respondents: Questionnaires sent by e-mail can help to
survey the opinion of the people living in far-flung places.
Precision of the hypothesis: Appropriateness of the questionnaire also
depends upon how realistic is the hypothesis in the mind of the researcher.
The researcher must frame his/her questions in such a manner that they elicit
responses needed to verify the hypothesis.
Types of Questions
There are many types of questions that can be asked, but the way to get to the
correct answer is to know which the right question is. It requires knowledge and
expertise to design the correct type of questionnaire.
Self-Instructional
Material 113
Research Tools The following is a list of the different types of questions which can be included
in questionnaire design:
Open format questions: Open format questions are those which give the
respondent a chance to communicate their individual opinions. There are
NOTES no set answers to choose from. Responses from open format questionnaires
are insightful and even unexpected. Qualitative questions are an example of
open format questions. An ideal questionnaire is one which ends with an
open format question giving the respondents the chance to state their opinion
or ask for their suggestions.
Example: ‘State your opinion about the grading system in education.’
A respondent’s answer to an open-ended question is coded into a response
scale afterwards. An example of an open-ended question is a question where
the testee has to complete a sentence (sentence completion item).
Closed format questions: Multiple choice questions are the best example
of closed format questions. Closed format questions generate responses
that can be statistics or percentages in nature. Preliminary analysis can also
be performed with ease. Closed format questions have the added advantage
of being able to monitor opinions over a period of time as they can be put to
different groups at different intervals.
Example: ‘Who is not an educationist among the following?’
(i) Prof Yashpal, (ii) John Dewy, (iii) Milkha Singh, (iv) Rabindranath Tagore.
Leading questions: These types of questions force the audience to give a
particular type of answer.
Example: ‘How would you rate the grading system in India?’
(i) Fair, (ii) Good, (iii) Excellent and (iv) Superb
Likert questions: Likert questions can help you ascertain how strongly
your respondent agrees with a particular statement. Likert questions can
also help to assess liking and disliking.
Example: ‘Are you punctual in attending your classes?’
(i) Always, (ii) Mostly, (iii) Normally, (iv) Sometimes, (v) Never
Rating scale questions: In rating scale questions, the respondent is asked
to rate a particular issue on a scale that may range from poor to good.
Rating scale questions usually have an even number of choices, so that
respondents are not given the choice of a middle option.
Example: ‘How was the food at the restaurant?’
(i) Good, (ii) Fair, (iii) Poor (iv) Very Poor
Questions to be avoided during preparation of a questionnaire
The following questions should be avoided when preparing a questionnaire:
Embarrassing questions: Embarrassing questions are those that ask
respondents about their personal and private life. Embarrassing questions
Self-Instructional are mostly avoided.
114 Material
Positive/Negative connotation questions: While defining a question, Research Tools
Self-Instructional
116 Material
Use another external criterion like consultation of documents or interview Research Tools
Self-Instructional
Material 117
Research Tools The way the data has been compiled will determine what information can
be gathered, e.g., if the response option is yes/no then you will only know
how many or what per cent of your sample answered yes/no, we will not
know how the average respondent answered.
NOTES
The questions asked (closed, multiple-choice, and open) should adhere to
the statistical data analysis techniques available and your goals.
Questions and prepared responses to choose from should not be biased.
A biased question or questionnaire influences the responses given.
The order in which the questions are presented or asked is also important
as the earlier questions and their responses may influence the later ones.
The wording should be kept simple to avoid ambiguity. Ambiguous words
may cause misunderstanding, possibly invalidating questionnaire results.
Double negatives should also be avoided.
Questions should address only one issue at a time so that the respondent is
not confused as to what response is required.
The list of possible responses should be comprehensive so that respondents
should not find themselves without a suitable response. A solution to this
would be to add the category of ‘other’.
Categories in the questionnaire should be kept separate. For example, in
both the ‘married’ category and the ‘single’ category—there may be a need
for separate questions on marital status and living situation.
Writing style should be informal yet to the point and suitable for the target
audience.
Personal questions about age, income, marital status, etc., should be placed
at the end of the survey so that even if the respondent is hesitant to give out
personal information, they would still have answered the other questions.
Questions which try to trick the respondent may end in inaccurate responses.
Presentation which is pleasing to the eye with the use of colours and images
can end up distracting the respondent.
Numbering the questions would be helpful.
Whoever administers the questionnaire, be it research staff, volunteers or
whether self-administered by the respondents, it should have clear, detailed
instructions.
Factors affecting reliability of answers
Factors affecting reliability of answers are as follows:
Confusing questions: If the questions are not easily intelligible or they are
capable of being interpreted in more than one way, the answers are unreliable,
because the answer may be the result of misinterpretation of the questions
not intended by the researcher.
Self-Instructional
118 Material
Prejudice regarding sample: The responses received from the sample Research Tools
5.4 SUMMARY
Self-Instructional
126 Material
There are different types of interviews such as group interview, diagnostic Research Tools
Self-Instructional
Material 127
Research Tools Questionnaire: It refers to a tool for research comprising a list of questions
whose answers provide information about the target group, individual or
event.
Rating scale: It refers to a scale used to evaluate the personal and social
NOTES
conduct of a learner.
Research proposal: It refers to a document proposing a research project,
generally in the sciences or academia, and generally constitutes a request
for sponsorship of that research.
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Self-Instructional
128 Material
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils, Research Tools
Self-Instructional
Material 129
Descriptive Research
6.0 INTRODUCTION
The case study method is mainly used for the purpose of qualitative analysis. It
involves a thorough and complete examination of a social unit. A social unit can
either be a person, a family, an institution, a cultural group or even the entire
community. A case study involves the in-depth study of a particular subject. The
case study method emphasizes on the complete investigation of only restricted
number of events related to a subject and the relationship between the different
events. The main objective of a case study is to determine the factors that are
responsible for the behaviour patterns of the given unit in totality.
In the words of the eminent researcher [Link], ‘The case study method is
a technique by which individual factor, whether it be an individual or just an
episode in the life of an individual or a group, is analysed in its relationship to
any other in the group’. Thus, a fairly exhaustive study of a person or group is
known as life or case study. Burgess has used the words, the ‘social microscope’ for
the case study method. Another researcher, Pauline V. Young, has defined the concept
of case study as ‘A comprehensive study of a social unit of a person, a group, a
social institution, a district, or a community’. In short, case study is a method
that involves qualitative analysis of an individual or a situation or an institution.
Characteristics of the Case Study Method
The following are certain characteristics of the case study method:
In this method, the researcher is allowed to take one or more than one
social unit for study. Instead of a social unit, the researcher can also select a
situation for study.
This method involves intensive study of the selected unit. As each unit
is studied for its minute details, the study takes a long period of time. This
Self-Instructional
Material 133
Descriptive Research help and ensures the correctness of the information collected about a
particular unit.
This method helps to determine the complex factors of a particular unit.
NOTES It also helps to determine the integrity of the selected unit with the other units.
This method follows the qualitative approach rather than the quantitative
approach.
In the case study method, efforts are made to determine the mutual inter-
relationship of the causal factors.
Evolution and Scope of the Case Study Method
In the field of sociology, the case study method is an extensively used research
technique. Frederic le Play introduced this method in the field of social investigation.
Herbert Spencer was the first to make use of case materials in his comparative
study of different cultures. This method is also used by anthropologists, historians,
novelists and dramatists to solve their problems related to their areas of interest.
Even management experts use this method to obtain clues of certain management
problems. Conclusively, the case study method is used in different disciplines.
Major Phases of the Case Study Method
The major phases involved in case study method are as follows:
Identification and resolution of the status of the phenomenon or the unit to
be examined.
Accumulation of data and selection of the phenomenon.
Investigation of the history of the selected phenomenon.
Analysis and recognition of informal factors related to the selected
phenomenon.
Application of corrective measures.
Review of programme to identify the effectiveness of the treatment applied.
Advantages of the Case Study Method
Some of the important advantages of the case study method are as follows:
As the case study method involves exhaustive study of a particular unit, the
complete behaviour pattern of the concerned unit is understood. According
to Charles Horton Cooley, case study deepens our perception and gives us
a clearer insight into life. It gets at behaviour directly and not by an indirect
and abstract approach.
With the help of a case study, a researcher can obtain genuine and progressive
record of personal experiences.
It helps a researcher to determine the natural history of the selected unit. It
also helps determine the relationship of the selected unit with the social
factors of the surrounding environment.
Self-Instructional
134 Material
It also helps to frame relevant hypotheses along with the data, which may Descriptive Research
Self-Instructional
Material 135
Descriptive Research
6.4 ETHNOGRAPHY
opposed to unified, cohesive, fixed and static. It should also be noted that while
cultures carry the ‘different-but-equal’ view, critical ethnography openly assumes
that cultures are not positioned equally in power relations. Further, critical
ethnography assumes that the descriptions of culture are shaped by the biases of NOTES
the researcher, the project sponsors, the audience, or the dominant communities.
Hence, cultural representations are deemed to be partial and partisan. Studies that
adopt the ethnographic approach should be conducted against the backdrop of
the theoretical assumptions behind this research initiative.
Data
Provides evidence of cohabitating or spending considerable time with people
who are in the study setting, by observing and recording their activities as
they unfolded through notes or journals, (Emerson, Fretz and Shaw, 1995),
audio and video recordings, or both. One of the trademarks of ethnography
is the extended and first-hand participant observations of their interactions
with participants in the study setting.
Records participants’ beliefs as well as their attitudes through typical means
such as notes or transcriptions of informal conversation and interviews, as
well as participant journals (Salzman, 2001).
Includes multiple sources of data. Besides observation and interactions with
participants, these sources can include life histories (Darnell, 2001) or
narrations (Cortazzi, 2001), photography, audio or video recordings
(Nastasi, 1999), written documents (Brewer, 2000), data that describes
historical trends, as well as questionnaires and surveys (Salzman, 2001).
Often called for in critical ethnography (and also in several cases of
descriptive or interpretive ethnography), to use additional sources of data
and reflection including:
o Evidence to show how the differences in power between you and the
informants or subjects were addressed. It is idealistic to assume that
differences in power may be totally eliminated, and hence what must
be addressed is how these differences were managed, amended, or
moved and also the influence that they had on the data gathered.
o The attitudes as well as biases towards the community and its culture.
There needs to be a record of how perspectives got modified as the
research progressed and how these modifications impacted the data
that was collected.
o The impact that your behaviour and activities have had on the
community. One must state if one was personally involved in the ethical,
social, or political challenges faced by the community. The data should
also contain the manner in which this involvement could have provided
deeper insights or impacted the research (and also the manner in which
the tensions were addressed).
Self-Instructional
Material 137
Descriptive Research o Expose the contradictions in the statements made by the insiders or
informants. Rather than opting for a particular data set over another,
or attempting to tie up all the loose ends in order to arrive at
generalizations, one should wade through the diverse insider
NOTES perspectives in order to more accurately represent the complex nature
of the culture.
o A wider insight and understanding of the context within which the
culture prevails. Context creation is a continuous activity taking place
even as the informants are interacting with the researcher. However,
the data must expose the manner in which external forces outside the
community shape culture. A study of the manner in which local culture
is shaped by social and political institutions and also obtains pre- and
post-research data on the status of the culture.
Analysis and Interpretation of Data
The emic perspective addresses the attitudes, beliefs, behaviours and practices of
the participants. This assists in achieving the objective of ethnography which is to
develop a comprehensive understanding of how people embedded in specific
contexts experience and react to their social and cultural worlds.
Ethic perspective refers to situations in which the researcher approaches
outsiders to analyse various behaviours or phenomena related with the
culture under study.
Symbols refer to any material such as architecture technology as a source
of information. Ethnographic researcher uses these symbols in understanding
the participant’s behaviour.
Tacit knowledge refers to deep and hidden information about cultural beliefs
and assumptions, but this knowledge is never formally or informally discussed
with the participants. Researchers use this knowledge individually.
Practice reflexivity is an introspective process of self-examination and
self-disclosure of one’s own background, identity or subjectivity, as well as
assumptions that are made and which may determine biases in data collection
and in its interpretation.
Approach is the data discovery and analysis in a manner that is inductive
and recursive. Be aware that patterns, categories and themes will evolve as
data collection progresses and be cautious not to impose these up front.
Evidence of triangulation should be depicted in the report. It is the
systematic process of scanning multiple data sources of information and
deriving conclusions so as to confirm or disregard evidence.
Specific context and particular time period are important features of
ethnographic knowledge. This is because of its first-hand and experiential
nature. However, most modern ethnographers acknowledge that the cultures
Self-Instructional
138 Material
being studied are unstable and ever-evolving. They pay attention to exploring Descriptive Research
6.7 SUMMARY
Self-Instructional
Material 141
Descriptive Research
6.8 KEY WORDS
Short-Answer Questions
1. Define the term ‘descriptive research’.
2. Write a short note on correlation and prediction studies.
3. List the limitations of correlation studies.
Long-Answer Questions
1. ‘A case study involves the in-depth study of a particular subject.’ Explain
the statement.
2. Examine the ethnographic method of descriptive research.
3. Critically analyse the analytical research method.
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
142 Material
Historical Research
7.0 INTRODUCTION
7.1 OBJECTIVES
Self-Instructional
Material 143
Historical Research
7.2 MEANING AND SCOPE OF HISTORICAL
RESEARCH
Self-Instructional
Material 145
Historical Research Studying the History of Institutions and Organizations: While studying
such history, the same general method applies as for the study of a University.
For example, one may study the history of the growth and development of
National Law Universities, IIMs, etc.
NOTES
7.2.3 Advantages and Disadvantages of Historical Research
The advantages of historical research are:
The researcher is not physically involved in the situation under study.
No danger of experimenter-subject interaction.
Documents are located by the researcher, data is gathered, and conclusions
are drawn out of sight.
Historical method is much more synthetic and eclectic in its approach than
other research methods, using concepts and conclusions from many other
disciplines to explore the historical record and to test the conclusions arrived
at by other methodologies.
Perhaps more than any other research method, historical research provides
librarians with a context. It helps to establish the context in which librarians
carry out their work. Understanding the context can enable them to fulfil
their functions in society.
It provides evidence of ongoing trends and problems.
It provides a comprehensive picture of historical trends.
It uses existing information.
Historical research suffers from several limitations, some are natural due to the
very nature of the subject and others extraneous to it and concerning the capabilities
of the researcher.
Good historical research is not lazy. It is slow, painstaking and exacting. An
average researcher finds it difficult to cope with these requirements.
Historical research requires a high level of knowledge, language skills and
art of writing on the part of the researcher.
Historical research requires a great commitment to methodological scholarly
activity.
Sources of data in historical researches are not available for the direct use
of the researcher and historical evidence is, by and large, incomplete.
Interpretation of data is very complex.
Through historical research, it is difficult to predict the future.
Scientific method cannot be applied to historical evidence.
Modern electronic aids (like computers) have not contributed much towards
historical research.
Self-Instructional
146 Material
It is not possible to construct ‘historical laws’ and ‘historical theories’. Historical Research
Man is more concerned with the present and future and has a tendency to
ignore the past.
Time-consuming. NOTES
Resources are scarce.
Data can be contradictory.
The research may not be conclusive.
Gaps in data cannot be filled as there are no additional sources of
information.
A historian can generalize but not predict or anticipate, can take precautions
but not control; can talk of possibilities but not probabilities.
7.2.4 Process of Historical Research
Historical research includes the delimitation of a problem, formulating hypothesis
or tentative generalizations, gathering and analysing data, and arriving at conclusions
or generalizations based upon deductive-inductive reasoning. However, according
to Ary, et al., (1972) the historian lacks control over both treatment and
measurement of data. He has relatively little control over sampling and he has no
opportunity for replication. As historical data is the closed class of data located
along a fixed temporal locus, the historian has no choice of sampling his data. He
is supposed to include every type of data that comes his way. Historical research
is not based upon experimentation, but upon reports of observation, which cannot
be authenticated. The historian handles data which are mainly traces of past events
in the form of various types of documents, relics, records and artefacts, which
have a direct or indirect impact on the event under study.
In deriving the truth from historical evidence, the major difficulty lies in the
fact that the data on which historical research is based are relatively inadequate.
It may be difficult to determine the data of occurrence of a certain historical
event partly because of changes brought out in the system of calendar and partly
due to incomplete information.
Historical research attempts to establish facts to arrive at a conclusion
concerning the past events.
Steps in Historical Research: The steps involved in undertaking a historical
research are not different from other forms of research. But the nature of the
subject matter presents a researcher with some peculiar standards and techniques.
In general, historical research involves the following steps:
Self-Instructional
Material 147
Historical Research
NOTES
Step 1: The first step is to make sure the subject falls in the area of the history of
education. One topic could be the study of the various educational systems and
how they have changed with the passing of time. On the other hand, studying
‘contributions of education’ as a component of national history can be of interest
to a researcher. The researcher may be interested in a historical investigation of
those aspects of education that have not been touched upon by any studies yet.
Moreover, the researcher may be interested in re-examining the validity of current
interpretations of certain historical problems which have already been studied.
Step 2: This necessitates that a thought is given to the various aspects of the
problem and various dimensions of the problem are identified. Hypothesis also
needs to be formulated. The hypothesis in historical research may not be able to
be tested, they are written as explicit statements that tentatively explain the
occurrence of events and conditions. While formatting a hypothesis, a researcher
may formulate questions that are most appropriate for the past events he is
investigating. Research is then directed towards seeking answers to these questions
with the help of the evidence.
Self-Instructional
148 Material
Step 3: Collection of historical evidence involves following two sub-steps: Historical Research
External criticism is also known as lower criticism. It involves testing the sources
of data for integrity, i.e., every researcher must test the information received to
ensure that any source of data is in fact what it seems to be. External criticism NOTES
helps to determine whether it is what appears or claims to be and whether it reads
true to the original so as to save the researcher from being the victim of fraud. On
the whole, the general criteria followed for such criticism depends on:
A good chronological sense, a versatile intellect, common sense, an intelligent
understanding of human behaviour, and plenty of patience and persistence
on the part of the researcher.
Recent validation of the quality of the source.
A good track record of the source.
This information may be found in relevant literature. Thereafter, these literary
sources can be verified for genuineness of content by verifying signatures,
handwriting, writing styles, language, etc. Further, material sources of information
can be verified through physical and chemical tests on the ink, paint, paper, cloth,
metal, wood, etc.
(b) Internal Criticism
After the integrity of the data sources are established, the actual data content is
subject to verification—this process is known as the internal criticism of the data.
It is also called higher criticism which is concerned with the validity, truthfulness, or
worth of the content of document.
At the outset, the information obtained through a particular source is
examined for internal consistency. The higher the internal consistency, the greater
the accuracy. The researcher should establish the literal as well as the real meaning
of the content within its historical context.
This is followed by an evaluation of the external consistency of the data.
This is important because, although the authorship of a report is established, the
report may comprise distorted pictures of the past. For verifying that the content is
accurate, the researcher should. firstly compare the information received through
two independent sources, and secondly match new information obtained with the
information already on hand which has been tested for reliability. Fox (1969)
suggested three major principles that need to be followed in order to establish
external consistency of the data: (i) Data from two independent sources to be
matched for consistency, (ii) Data must have been obtained from at least one
independent primary source, and (iii) Data should not be gathered from a source
that has a track record of providing contradictory information. It is recommended
that the researcher apply his professional knowledge and judgment to make a final
evaluation in case it is not possible to find matching information from two comparable
sources.
Self-Instructional
Material 151
Historical Research The following series of questions have been listed by Good, Barr and Scates
(1941) to guide a researcher in the process of external and internal criticism of
historical data:
Who was the author, not merely what his name was but what his personality,
NOTES
character and position were like, etc.?
What were his general qualifications as a reporter—alertness, character
and bias?
What were his special qualifications as a reporter of the matters here treated?
How was he interested in the events related?
Under what circumstances was he observing the events?
Had he the necessary general and technical knowledge for learning and
reporting the events?
How soon after the events was the document written?
How was the document written, from memory, after consultation with others,
after checking the facts, or by combining earlier trial drafts?
How is the document related to other documents?
Is the document an original source—wholly or in part? If the latter, what
parts are original, what borrowed? How credible are the borrowed
materials? How accurately is the borrowing done? How is the borrowed
material changed and used?
Perpetually, the researcher needs answers for all these questions and,
therefore, he has to depend, somewhat, upon evidence he can no longer verify. At
times, he will have to rely on the inferences based upon logical deductions in order
to bridge the gaps in the information.
7.2.7 Purpose of Historical Research
Historical research is carried out to serve the following purposes:
To discover the context of an organizational situation: In order to
explore and explain the past, a historian aims to seek the context of an
organization/a movement/ the situation being studied.
To answer questions about the past: There are many questions about
the past to which we would like to find answers. Knowing the answers can
enable us to develop an understanding of past events.
To study the relationship of cause and effect: There is a cause and
effect relationship between two events. A historian would like to determine
such a relationship.
To study the relationship between the past and the present: The past
can often help us get a better perspective about current events. Thus, a
researcher aims to identify the relationship between the past and the present,
Self-Instructional whereby we can get a clear perspective of the present.
152 Material
To reorganize the past: A historian reconstructs the past systematically Historical Research
Self-Instructional
154 Material
Historical Research
7.3 ANSWERS TO CHECK YOUR PROGRESS
QUESTIONS
7.4 SUMMARY
Self-Instructional
Material 155
Historical Research the data. It is also called higher criticism which is concerned with the validity,
truthfulness, or worth of the content of document.
Historians are greatly interested in recording and evaluating the
accomplishments of leading individuals and different kinds of organizations
NOTES
including institutions and agencies as these influence historical events.
Historical synthesis and interpretation are considered an art, which is
subjective in nature. This raises a serious problem of subjectivity. ‘Historical
synthesis is necessarily a highly subjective art. It involves the intuitive
perception of patterns and relationships in the complex web of events, as
well as the art of narrative writing.
Short-Answer Questions
1. Write a short note on the meaning and scope of historical research.
2. List the types of historical research.
3. Briefly mention the sources of data in historical research.
4. Mention the purposes of conducting historical research.
Long-Answer Questions
1. Discuss the advantages and disadvantages of historical research.
2. Explain the process of historical research.
3. How is historical data evaluated?
Self-Instructional
156 Material
Historical Research
7.7 FURTHER READINGS
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall. NOTES
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
Material 157
Experimental Research
UNIT 8 EXPERIMENTAL
RESEARCH
NOTES
Structure
8.0 Introduction
8.1 Objectives
8.2 Design: Pre-Experimental, Quasi-Experimental and True-Experimental
8.3 Internal and External Experimental Validity
8.4 Answers to Check Your Progress Questions
8.5 Summary
8.6 Key Words
8.7 Self Assessment Questions and Exercises
8.8 Further Readings
8.0 INTRODUCTION
Experiments are used to infer causality where the researcher actively manipulates
one or more causal variables and measure their effects on the dependent variable.
There are three necessary conditions for inferring causality: (i) concomitant variation
(ii) time order of occurrence of variables, and (iii) the absence of other possible
causal factors. Various concepts like independent variables (treatments), test units,
dependent variables, exogenous variables are used in conducting an experiment.
An experiment can be conducted under different environmental conditions, namely,
laboratory and field. The researcher has two goals while conducting an experiment:
(i) to keep the internal validity of the experiment very high and (ii) to make
generalization of the results of the experiments to a wider population.
In this unit, you will learn the different experimental research designs. This
will include a discussion on the internal and external validity of experimental research.
8.1 OBJECTIVES
Self-Instructional
158 Material
Experimental Research
8.2 DESIGN: PRE-EXPERIMENTAL, QUASI-
EXPERIMENTAL AND TRUE-EXPERIMENTAL
Self-Instructional
Material 159
Experimental Research would make the design invalid for making any causal inferences on account
of the following reasons:
The economic condition might have changed during the two periods
(history).
NOTES
The test units may mature over time (maturation).
The pre-test measurement on the test units may influence the
performance (testing).
The prices of goods might have changed over time (instrumentation).
Test units might not have been selected at random (selection bias).
Some test units might have left before the experiment was complete
(mortality).
Test units might be self-selected on the basis of the current poor
performance and may have a better period ahead because of sheer
luck (regression).
(iii) Static Group Comparison: This design is symbolically written as:
Group 1 - X O1
Group 2 - O2
This design uses two treatment groups. Test units in both the groups are not
selected at random. The first group, called the experimental group, is
subjected to the treatment X, whereas the second group, namely, the control
group, is not subjected to any treatment. Both groups are measured only
after the treatment has been presented. Thus, it is critical to understand that
in this design the exposure as well as the experimental treatment is not
under the control of the researcher. Consider the following example:
A study wants to assess the relationship of ‘family support’ (measured by
the presence of domestic help or spouse/family’s help in carrying out domestic
chores) with the work–life balance of BPO women employees. Here, the
presence or absence of help is ascertained and then we can measure the
work–life balance. Thus the design is essentially ex-post facto and any
segregation into experimental or control group is not made by the researcher.
The treatment effect could be measured by O1 – O2. However, this difference
could be attributed to at least selection bias and mortality. Moreover, since the
test units are not selected at random, the two groups could differ prior to the
application of treatment. All these are sufficient to make the design invalid for
drawing any causal inferences.
Quasi-Experimental Designs
In quasi-experimental design the researcher can control when measurements are
taken and on whom they are taken. However, this design lacks complete control
of scheduling of treatment and also lacks the ability to randomize test units’ exposure
Self-Instructional
160 Material
to treatments. As the experimental control is lacking, the possibility of getting Experimental Research
Self-Instructional
Material 161
Experimental Research
NOTES
taken after the training programme. The treatment effect (sales training) is
found by comparing the average sales of the two groups before and after
the training programme. The major drawback of this design is the possibility
of the interactive effect in the experimental group. NOTES
Self-Instructional
164 Material
of difference of various post-test and pre-test measurement would give the Experimental Research
following results:
Experimental Group 1:
O2 – O1 = Treatment effect + Extraneous factors without NOTES
interactive testing effect + Interactive testing effect ...(i)
Control Group 1:
O4 – O3 = Extraneous factors without interactive testing effect ...(ii)
As this group was not subjected to any treatment, there would not be any
interactive testing effect.
Experimental Group 2:
O5 – O1 = Treatment effect + Extraneous factors without
interactive testing effect ...(iii)
O5 – O3 = Treatment effect + Extraneous factors without
testing effect ...(iv)
As there was actually no pre-test measurement, the interactive testing effect
cannot occur here.
Control Group 2:
O6 – O1 = (Extraneous factors without testing effect) ...(v)
O6 – O3 = (Extraneous factors without testing effect) ...(vi)
As the group was not subjected to any treatment, the difference in
measurement would only indicate the effect of extraneous factors without
interactive testing effect.
By taking the average of (v) and (vi), one gets:
O1 O3
O6 = (Extraneous factors without testing effect) ...(vii)
2
By taking the average of (iii) and (iv), one obtains:
O1 O3
O5 = Treatment effect + Extraneous factors without
2
testing effect ...(viii)
By subtracting (vii) from (viii), one obtains:
O1 O3 O1 O3
O5 O6 = O5 – O6 = Treatment effect
2 2
By subtracting (viii) from (i), one obtains:
O O3
O2 O1 O5 1 = Interacting testing effect
2
Self-Instructional
Material 165
Experimental Research Therefore, this design has helped not only in measuring the effect of treatment,
but also in obtaining magnitude of the interactive testing effect and extraneous
factors.
To conduct this experimental design, the time and cost required are enormous
NOTES
and therefore, this design is not commonly used in research. However, as
seen, this experimental design guarantees the maximum internal validity. In
businesses where establishing cause-and-effect relationship is very crucial
for survival, this design is useful.
Statistical Designs
Statistical designs allow for statistical control and analysis of external variables.
The main advantages of statistical design are the following:
The effect of more than one level of independent variable on the dependent
variable can be manipulated.
The effect of more than one independent variable can be examined.
The effect of specific extraneous variable can be controlled.
Included in this category are the following designs:
(i) Completely Randomized Design: This design is used when a researcher
is investigating the effect of one independent variable on the dependent
variable. The independent variable is required to be measured in nominal
scale i.e. it should have a number of categories. Each of the categories of
the independent variable is considered as the treatment. The basic assumption
of this design is that there are no differences in the test units. All the test units
are treated alike and randomly assigned to the test groups. This means that
there are no extraneous variables that could influence the outcome.
Suppose we know that the sales of a product is influenced by the price
level. In this case, sales are a dependent variable and the price is the
independent variable. Let there be three levels of price, namely, low, medium
and high. We wish to determine the most effective price level i.e. at which
price level the sale is highest. Here the test units are the stores which are
randomly assigned to the three treatment level. The average sales for each
price level is computed and examined to see whether there is any significant
difference in the sale at various price levels. The statistical technique to test
for such a difference is called ANalysis Of VAriance (ANOVA).
This design suffers from the main limitation that it does not take into account
the effect of extraneous variables on the dependent variable. The possible
extraneous variables in the present example could be the size of the store,
the competitor’s price and price of the substitute product in question. This
design assumes that all the extraneous factors have the same influence on all
Self-Instructional
166 Material
the test units which may not be true in reality. This design is very simple and Experimental Research
inexpensive to conduct.
(ii) Randomized Block Design: As discussed, the main limitation of the complete
randomized design is that all extraneous variables were assumed to be
NOTES
constant over all the treatment groups. This may not be true. There may be
extraneous variables influencing the dependent variable. In the randomized
block design it is possible to separate the influence of one extraneous variable
on a particular dependent variable, thereby providing a clear picture of the
impact of treatment on test units.
In the example considered in the completely randomized design, the price
level (low, medium and high) was considered as an independent variable
and all the test units (stores) were assumed to be more or less equal.
However, all stores may not be of the same size and, therefore, can be
classified as small, medium and large size stores. In this design, the extraneous
variable, like the size of the store could be treated as different blocks. Now
the treatments are randomly assigned to the blocks in such a way so that
each treatment appears in each block at least once. The purpose of forming
these blocks is that it is hoped that the scores of the test units within each
block would be more or less homogeneous when the treatment is absent.
What is assumed here is that block (size of the store) is correlated with the
dependent variable (sales). It may be noted that blocking is done prior to
the application of the treatment.
In this experiment one might randomly assign 12 small-sized stores to three
price levels in such a way that there are four stores for each of the three
price levels. Similarly, 12 medium-sized stores and 12 large-sized stores
may be randomly assigned to three price levels. Now the technique of analysis
of variance could be employed to analyse the effect of treatment on the
dependent variable and to separate out the influence of extraneous variable
(size of store) from the experiment.
(iii) Latin Square Design: This design is employed when the researcher is
interested in separating out the influence of two extraneous variables.
Suppose the interest is to study the influence of price (treatment) on sales.
Let there be three levels of price categorizes, namely, low (X1), medium
(X2) and high (X3). The sales could be influenced by two extraneous
variables, namely, store size and type of packaging. For the application of
the Latin square design, the number of categories of two extraneous variables
should be equal to the number of levels of treatments. This is a necessary
condition for the use of Latin square design. The store could be of size –
small (1), medium (2) and large (3) and type of packaging could be I, II and
III. The Table 8.1 below presents the layout of the Latin square design.
Self-Instructional
Material 167
Experimental Research Table 8.1 Latin Square Design for Various Levels of Price
I II III
NOTES 1 (Small) X1 X2 X3
2 (Medium) X2 X3 X1
3 (Large) X3 X1 X2
It may be noted that the rows and columns represent those extraneous
variables whose effect is to be controlled and measured. There are three
categories of row variable (size of store) and three categories of column
variable (type of packaging). This would result in 3 × 3 Latin square.
One point that has to be kept in mind is that the treatment should be assigned
randomly to cells in such a way that each treatment occurs once and only
once in each row and in each column. The treatments exhibited in Table 8.2
satisfy this condition.
Use of this design helps to measure statistically the effect of a treatment on
the dependent variable and also the measurement of an error resulting from
two extraneous variables. This design, indeed has a very complex setup
and is quite expensive to execute.
(iv) Factorial Design: A factorial design may be employed to measure the
effect of two or more independent variables at various levels. The factorial
designs allow for interaction between the variables. An interaction is said to
take place when the simultaneous effect of two or more variables is different
from the sum of their individual effects. An individual may have a high
preference for mangoes and may also like ice-cream, which does not mean
that he would like mango ice cream, leading to an interaction.
The sales of a product may be influenced by two factors, namely, price
level and store size. There may be three levels of price—low (A1), medium
(A2) and high (A3). The store size could be categorized into small (B1) and
big (B2). This could be conceptualized as a two-factor design with
information reported in the form of a table. In the table, each level of one
factor may be presented as a row and each level of another variable would
be presented as a column. This example could be summarized in the form
of a table having three rows and two columns. This would require 3 × 2 =
6 cells. Therefore, six different level of treatment combinations would be
produced each with a specific level of price and store size. The respondents
would be randomly selected and randomly assigned to the six cells. The
tabular presentation of 3 × 2 factorial design is given in Table 8.2.
Self-Instructional
168 Material
Table 8.2 3 × 2 Factorial Design for Price Level and Store Size Experimental Research
Price Store
Small (B1) Big (B2)
Low Level (A1) A1B1 A1B2 NOTES
Medium Level (A2) A2B1 A2B2
High Level (A3) A3B1 A3B2
Self-Instructional
Material 169
Experimental Research designs the researcher has control over when the measurements are to be
taken and on whom they are taken. However, the design lacks complete
control of scheduling of treatment and also lacks ability to randomize test
units exposure to treatments. Included in the category of true-experimental
NOTES design are (i) pre-test–post-test control group, (ii) post-test–only control
group and (iii) Solomon four-group design. In these designs, the researcher
can randomly assign test units and treatments to experimental groups. The
researcher is able to eliminate the effect of extraneous variables from both
control and experimental groups. The statistical designs covered here are
(i) completely randomized design, (ii) randomized block design, (iii) Latin
square design, and (iv) factorial design. The statistical designs help to
(i) study the effect of more than one level of independent variables on the
dependent variable; (ii) study the effect of more than one independent
variable and (iii) the effect of specific extraneous variables.
Independent and Repeated Measures designs
In the independent measure design, different set of participants are used for each
condition of research. The advantage of this type of research design is that the
problem of order of conditions, where participants behave differently based on
the order is eliminated. The disadvantage is that the individual differences of groups
may create error in the research. This type of research design is called between
groups.
In repeated measures design, the same group of individuals are tested for
different conditions. The advantage is that the problem of individual difference
between groups is eliminated. The disadvantage is that a fewer group is required
for this type of research, there may be a problem of order effects and the range of
use is limited as participants are repeated. This type of research design is also
called within group design.
Nested and Single-subject designs
Nesting design
Nested design is a research design in which levels of one factor are hierarchically
subsumed under or nested within levels of another factor.
Crossed design is a research design that has at least two factors that are
crossed, i.e. every category of one factor co-occurs in the design with every
category of the other factor
Single subject design
In design of experiments, single-subject design or single-case research design is
a research design most often used in applied fields of psychology, education, and
human behaviour in which the subject serves as his/her own control, rather than
using another individual/group.
Self-Instructional
170 Material
The following are the requirements of single-subject designs: Experimental Research
In this section, you will learn about the concept of internal and external experimental
validity and how to control extraneous and intervening variables.
Internal Validity
Internal validity is considered as a property of scientific studies which indicates the
extent to which an underlying conclusion based on a study is warranted. This type
of warrant is constituted by the extent to which a study minimizes systematic error
or ‘bias’. If a causal relation between two variables is properly demonstrated then
the inferences are said to possess internal validity. A fundamental inference may be
based on a relation when the following three criteria are satisfied:
1. The ‘cause’ precedes the ‘effect’ in time (temporal precedence).
2. The ‘cause’ and the ‘effect’ are related (covariation).
Self-Instructional
Material 171
Experimental Research 3. There are no plausible alternative explanations for the observed covariation
(non-spuriousness).
Internal validity refers to the ability of a research design for providing an adequate
test of an hypothesis and the ability to rule out all plausible explanations for the
NOTES
results but the explanation being tested. For example, let us consider that a
researcher decides that a particular medication prevents the development of heart
disease because he found that research participants who took the medication
developed lower rates of heart disease than those who never took the medication.
This interpretation of the study’s results is likely to be correct, however, only if the
study has high internal validity. In order to have high internal validity, the research
design must have controlled the directionality and third-variable problems, as well
as for the effects of other extraneous variables. In short, the researcher would
have needed to perform an experimental study in which:
Participants were randomly assigned to the experimental and control groups.
Participants did not know whether they were taking the medication.
The most internally valid studies are experimental studies because they are
better than correlational and case studies at controlling for the directionality and
third-variable problems, as well as for the effects of other extraneous variables.
Threats to Internal Validity
The following are the various threats to internal validity:
Ambiguous Temporal Precedence: Lack of precision about the occurrence of
variable, i.e., which variable occurred first, may yield confusion that which variable
is the cause and which is the effect.
Confounding: Confounding is a major threat to the validity of fundamental
inferences. Changes in the dependent variable may rather be attributed to the
existence or variations in the degree of a third variable which is related to the
manipulated variable. Rival hypotheses to the original fundamental inference
hypothesis of the researcher may be developed where spurious relationships cannot
be ruled out.
Selection Bias: It refers to the problem that, at pre-test, differences between the
existing groups that may interact with the independent variable and thus be
‘responsible’ for the observed outcome. Researchers and participants bring to the
experiment a myriad of characteristics, some learned and others inherent. For
example, sex, weight, hair, eye, and skin color, personality, mental capabilities and
physical abilities, etc. Attitudes like motivation or willingness to participate can
also be involved. If an unequal number of test subjects have similar subject-related
variables during the selection step of the research study, then there is a threat to
the internal validity.
Repeated Testing: It is also referred to as testing effects. Repeatedly measuring
or testing the participants may lead to bias. Participants of the testing may remember
Self-Instructional
the correct answers or may be conditioned to know that they are being tested.
172 Material
Repeatedly performing the same or similar intelligence tests usually leads to score Experimental Research
gains instead of concluding that the underlying skills have changed for good. This
type of threat to internal validity provides good rival hypotheses.
Regression toward the Mean: When subjects are selected on the basis of
NOTES
extreme scores (one far away from the mean) during a test then this type of threat
occurs. For example, in a testing when children with the bad reading scores are
selected for participating in a reading course, improvements in the reading at the
end of the course might be due to regression toward the mean and not the course’s
effectiveness actually. If the children had been tested again before the course started,
they would likely have obtained better scores anyway.
External Validity
External validity is considered as the validity of generalized (causal or fundamental)
inferences in scientific studies. It is typically based on experiments as experimental
[Link] other words, it is the degree to which the outcomes of a study can be
generalized to other situations and people.
If inferences about cause and effect relationships which are based on a
particular scientific study may be generalized from the unique and characteristics
settings, procedures and participants to other populations and conditions then
they are said to possess external validity. Causal inferences possessing high degrees
of external validity can reasonably be expected to apply:
To the target population of the study, i.e., from which the sample was drawn.
It is also referred to as population validity.
To the universe of other populations, i.e., across time and space.
An experiment using human participants often employ small samples which
are obtained from a single geographic location or with characteristics features is
considered as the most common threat to external validity. Due to this reason, one
cannot be certain that the conclusions drawn about cause and effect relationships
do actually apply to people in other geographic locations or without these particular
features.
External validity refers to the ability of a research design for providing
outcomes that can be generalized to other situations, especially to real-life situations.
For instance, if the researcher in the hypothetical heart disease medication study
found that the medication, under controlled conditions, prevented the development
of heart disease in research participants, he would want to generalize these findings
to state that the medication will prevent heart disease in the general population.
However, let us consider that the research design required the elimination of many
potential participants, such as people who abuse alcohol or other drugs, suffer
from diabetes, weigh more than average for their height, and have never suffered
from a mood or anxiety disorder. These are common risk factors for heart disease
and, by eliminating these factors; the outcomes of the study would provide little
evidence that the medication will be effective for people with these risk factors. In
Self-Instructional
Material 173
Experimental Research other words, the study would have low external validity and, hence, its outcomes
to the general population could not be generalized.
This commonly happens in tests of antidepressant medications. Because researchers
want to make sure that the antidepressant effects of the medications being tested
NOTES
are not hidden by the effects of extraneous variables, they often have excluded
potential participants with one or more of the following characteristics:
People who are addicted to alcohol or illicit drugs.
People who take various medications.
People who have anxiety disorders (such as, phobic disorders).
People who suffer from depression with psychosis.
People with mild depression (because they would show only a small
response to the medication).
If a study excluded people with these characteristic features, then most of
the participants suffering from depression would be excluded from the final pool
of participants. The outcomes of the study, therefore, would provide little information
about how most depressed people will respond to the medication.
Threats to External Validity
A threat to external validity is an explanation of how you might be wrong in making
a generalization. Usually, generalization is limited when the cause, i.e., independent
variable depends on other factors; therefore, all threats to external validity interact
with the independent variable.
Aptitude-Treatment Interaction: The sample may have specific
characteristic features that may interact with the independent variable, limiting
generalization. For example, inferences based on comparative psychotherapy
studies often employ specific samples (e.g., volunteers, highly depressed,
no comorbidity). If psychotherapy is found effective for these sample
patients, will it also be effective for non-volunteers or the mildly depressed
or patients with concurrent other disorders?
Situation: All situational features, such as treatment conditions, time,
location, lighting, noise, treatment administration, investigator, timing, scope
and extent of measurement, etc. of a study potentially limit generalization.
Pre-Test Effects: If cause and effect relationships can only be found when
pre-tests are carried out, then this also limits the generality of the findings.
Post-Test Effects: If cause and effect relationships can only be found
when post-tests are carried out, then this also limits the generality of the
findings.
Reactivity (Placebo, Novelty and Hawthorne Effects): If cause and
effect relationships are found they might not be generalized to other situations
if the effects found only occurred as an effect of studying the situation.
Self-Instructional
174 Material
Rosenthal Effects: Inferences about cause-consequence relationships may Experimental Research
8.5 SUMMARY
(ii) randomized block design, (iii) Latin square design, and (iv) factorial
design.
The statistical designs help to (i) study the effect of more than one level of
NOTES
independent variables on the dependent variable; (ii) study the effect of
more than one independent variable and (iii) the effect of specific extraneous
variables.
Internal validity is considered as a property of scientific studies which indicates
the extent to which an underlying conclusion based on a study is warranted.
This type of warrant is constituted by the extent to which a study minimizes
systematic error or ‘bias’.
External validity is considered as the validity of generalized (causal or
fundamental) inferences in scientific studies. It is typically based on
experiments as experimental [Link] other words, it is the degree to which
the outcomes of a study can be generalized to other situations and people.
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
178 Material
Data Analysis
BLOCK - III
QUALITATIVE AND QUANTITATIVE DATA ANALYSIS
NOTES
UNIT 9 DATA ANALYSIS
Structure
9.0 Introduction
9.1 Objectives
9.2 Types of Measurement Scales
9.3 Descriptive and Inferential Data Analysis
9.3.1 Tools of Descriptive and Inferential Statistics
9.3.2 Quantitative, Qualitative, Parametric and Non-Parametric Analysis
9.4 Answers to Check Your Progress Questions
9.5 Summary
9.6 Key Words
9.7 Self Assessment Questions and Exercises
9.8 Further Readings
9.0 INTRODUCTION
In the research methodology, one of the major and significant tasks involved is
data analysis. This comes after the collection of data. Data analysis simply refers
to the process of interpreting the information collected, attaching significance to
the values and arriving at results. Data analysis in a sense is the stage where sense
is made of the data collected. Data analysis helps in giving explanation to the data
collected, allows comparison between different sets of data, assists in identification
of data outliers and helps in making predictions for the future. Data analysis requires
certain prerequisites to make it a smooth and consistent process. This requires the
selection of measurement scales, the type and method of analysis as well as the
classification of tests to be used for interpretation. All of these concepts will be
introduced in this unit.
9.1 OBJECTIVES
Self-Instructional
Material 179
Data Analysis
9.2 TYPES OF MEASUREMENT SCALES
There are four types of measurement scales—nominal, ordinal, interval and ratio
NOTES scales. We will discuss each one of them in detail. The choice of the measurement
scale has implications for the statistical technique to be used for data analysis.
Nominal scale: This is the lowest level of measurement. Here, numbers
are assigned for the purpose of identification of the objects. Any object which is
assigned a higher number is in no way superior to the one which is assigned a
lower number. In the nominal scale there is a strict one-to-one correspondence
between the numbers and the objects. Each number is assigned to only one object
and each object has only one number assigned to it. It may be noted that the
objects are divided into mutually exclusive and collectively exhaustive categories.
Examples of nominal scale:
What is your religion?
(a) Hinduism
(b) Sikhism
(c) Christianity
(d) Islam
(e) Any other, (please specify)
A Hindu may be assigned a number 1, a Sikh may be assigned a number 2,
a Christian may be assigned a number 3 and so on. Any religion which is assigned
a higher number is in no way superior to the one which is assigned a lower number.
The assignment of numbers is only for the purpose of identification. We also note
that all respondents have been divided into mutually exclusive and collectively
exhaustive categories. For example:
Are you married?
(a) Yes
(b) No
If a person is married, he or she may be assigned a number 101 and an
unmarried person may be assigned a number 102.
In which of the following departments do you work?
(a) Marketing
(b) HR
(c) Information Technology
(d) Operations
(e) Finance and Accounting
(f) Any other, (please specify)
Self-Instructional
180 Material
Here also, a person working for the marketing department may be assigned Data Analysis
a number 1, the one working for HR may be assigned a number 2 and so on.
Nominal scale measurements are used for identifying food habits (vegetarian
or non-vegetarian), gender (male/female), caste, respondents, brands, attributes, NOTES
stores, the players of a hockey team and so on.
The assigned numbers cannot be added, subtracted, multiplied or divided.
The only arithmetic operations that can be carried out are the count of each category.
Therefore, a frequency distribution table can be prepared for the nominal scale
variables and mode of the distribution can be worked out. One can also use chi-
square test and compute contingency coefficient using nominal scale variables.
Ordinal scale: This is the next higher level of measurement than the nominal
scale measurement. One of the limitations of the nominal scale measurements is
that we cannot say whether the assigned number to an object is higher or lower
than the one assigned to another option. The ordinal scale measurement takes
care of this limitation. An ordinal scale measurement tells whether an object has
more or less of characteristics than some other objects. However, it cannot answer
how much more or how much less. An ordinal scale tells us the relative positions
of the objects and not the difference between the magnitudes of the objects.
Suppose Shashi scores the highest marks in marketing and is ranked no. 1; Mohan
scores the second highest marks and is ranked no. 2; and Krishna scores third
highest marks and is ranked no. 3. However, from this statement we cannot say
whether the difference in the marks scored by Shashi and Mohan is the same as
between Mohan and Krishna. The only statement which can be made under ordinal
scale is that Shashi has scored higher than Mohan and Mohan has scored higher
than Krishna. The difference between the ranks does not have any meaningful
interpretation in the sense that it cannot tell the difference in absolute marks between
the three candidates. Another example of the ordinal scale could be the CAT
score given in percentile form. Suppose a candidate’s score is 95 percentile in the
CAT exam. What it means is that 95 per cent of the candidates that appeared in
the CAT examination have a score below this candidate, whereas only 5 per cent
have scored more than him. The actual score is how much less or more cannot be
known from this statement. Examples of the ordinal scale include quality ranking,
rankings of the teams in a tournament, ranking of preference for colours, soft
drinks, socio-economic class and occupational status, to mention a few. Some of
the examples of ordinal scales are listed below:
• Rank the following attributes while choosing a restaurant for dinner. The
most important attribute may be ranked one, the next important may be
assigned a rank of 2 and so on.
Self-Instructional
Material 181
Data Analysis
NOTES
Rank the following by placing a 1 beside the attribute you think is the most
important, a 2 beside the attribute you think is the second most important
and so on while purchasing a two-wheeler.
The interval scale data has an arbitrary origin (non-zero origin). The most
common example of the interval scale data is the relationship between Celsius and
Farenheit temperature. It is known that:
Self-Instructional
182 Material
Data Analysis
NOTES
This is of the form and b = /5/__/9/
and hence it represents the interval scale measurement. In the interval scale, the
difference in score has a meaningful interpretation while the ratio of the score on
this scale does not have a meaningful interpretation. This can be seen from the
following interval scale question:
How likely are you to buy a new designer carpet in the next six months?
Self-Instructional
Material 183
Data Analysis
NOTES
As the name suggests, descriptive statistics merely describe the data and consist
of methods and techniques used in collection, organization, presentation and analysis
Self-Instructional
Material 185
Data Analysis of data in order to describe the various features and characteristics of such data.
These methods can either be graphical or computational. Thus data can be
presented in the form of a chart or a table in order to show certain trends,
proportions, maximum and minimum values and so on. For example, if we simply
NOTES describe the number of workers in different types of industries in America, then
that would constitute descriptive statistics. In addition to the organization of data,
the field of descriptive statistics is concerned with the analysis of data so that the
data can be easily understood. Averages, proportions and other measures that
describe the spread of data around the average are also some of the measures
used to describe the data. By using these measures, we summarize the data and
even though we may lose the detail, we gain clarity and compactness. For example,
the following statistics, in their most summarized presentation describe in some
way the characteristics of the population from which they were drawn.
The ages of students in my statistics class range from 19 to 45 years.
The average IQ of students at our college is 140.
20 per cent of the students in my class are married.
All these examples simply summarize and describe the data. Not much can
be inferred from them, nor can definite decisions be made or conclusions drawn.
For a proper appreciation of the various descriptive statistics involved, it is
necessary to note that most of the statistical distribution have some common
features. Though the size of the variables varies from item to item, most of the
items are distributed in such a manner that if we move from the lowest value to the
highest value of the variable, the number of items at each successive stage increases
with a certain amount of regularity till we reach a maximum; and then as we proceed
further, they decrease with the similar regularity. If we plot the percentage frequency
density, i.e., the percentage of cases in an interval of unit variable width, we get
frequency curves of the type shown in Figure 9.1 (Note that the area under each
curve should be equal to 100, the total percentage points).
There are various ‘gross’ ways in which frequency curves can differ from
one another. Even when the ‘general’ shapes of the curves are the same (the area
under them already made equal by the strategy of plotting the percent density), the
details of the shape may change. Thus the curve B has a smaller spread than A, the
curve C is more peaky and curve E is less symmetrical. Even when the curves
have almost the same shape (i.e., same spread, peakiness, symmetry, etc.) as in
curves A and D, the two may differ in location along the variable axis. Thus the
items of distribution D are generally larger than those of A. So are those of B
compared to A. Thus, a kind of an ‘average’ location of the distribution along the
variable axis is an important descriptive statistics. These statistics are collectively
known as measures of location or of central tendency.
Self-Instructional
186 Material
Data Analysis
NOTES
Inferential Statistics
Inferential statistics can be defined as those methods that are used to estimate a
characteristic of a population or making of a decision concerning a population on
the basis of the results obtained from a sample taken from the same population.
The measured characteristics of the sample are known as sample statistics, while
the measured characteristics of the population are known as population parameters.
A major portion of statistics deals with making decisions, inferences, predictions
and forecasts about the population based on the results obtained from samples
taken from such populations.
The need for inferential statistical methods derives from the need for sampling.
As the population becomes large, it is usually too costly, too time consuming and
too cumbersome to take the entire population into consideration in order to obtain
our information of interest. Of course, the results obtained from the entire population
are the most accurate and if the population indeed is small, then it is advisable to
consider the entire population. However, when the population is large, sometimes
considered infinite, then sampling method is used.
The question is: How do these sample statistics relate to population
parameters? Can we state that the conclusions drawn from the analysis of the
sample are exactly the same as the conclusions that would be drawn from the
entire population from which the representative sample was taken? The answer is
unlikely. How close is the sample characteristics to the population characteristics
would depend upon the randomness of the sample as well as the size of the sample.
The more random the sample is and larger the sample is, the more closely its
characteristics would be with the population characteristics. This link, in terms of
the degree of closeness is provided by probability theory. Probability theory provides
the link by ascertaining the likelihood that the results from the sample reflect the
results from the population.
Our interest is not in finding the characteristics of a sample but our to find
the characteristics of the population. Sampling is simply a means to the end. For
Self-Instructional
Material 187
Data Analysis example, if we want to know the salary of university professors, we mean the
salary of all university professors and not simply of the sample we have taken.
Only then can observations and decisions be made in this regard. Similarly, if we
want to know what percentage of eligible voters will vote for Congress in the next
NOTES general elections in India, a sample in itself would not indicate that, and we cannot
ask the entire population. Our decisions and projections would be based on the
inclination of the entire population. A sample in itself would not mean much, if any
thing. How ever, if the sample truly represents the population, then we can draw
conclusions about the population on the basis of sample results. Appended to
these conclusions will be a probability statement specifying the likelihood or
confidence that the results from the sample reflect the voting behaviour of
the population. Usually, the margin of error is stated as plus or minus three to five
per cent.
Statistical inference deals with methods of inferring or drawing conclusions
about the characteristics of the population based upon the results of the sample
taken from the same population. The measured characteristics of the sample are
called sample statistics and the measured characteristics of the population are
known as population parameters. The question is: How do these sample statistics
relate to population parameters? Can we state that the conclusions drawn from
the analysis of the sample are exactly the same as the conclusions that would be
drawn from the entire population from which the representative sample was taken?
Following are some of the situations that the field of inferential statistics deals with.
Examples:
(a) Between 35 per cent and 40 per cent of graduate students in the universities
are married. These statistics refer to the entire population of graduate
students. It would be reasonable to assume that these percentages were
calculated on the basis of samples taken from the population of all graduate
students. The students in these samples were asked in order to know as to
how many of these students were married. The answers formed the basis
for drawing conclusions about the entire population of the graduate students.
(b) There is a definitive association between smoking and lung cancer.
This statement is the result of endless research on many samples taken and
studied in order to find out if there was any correlation between smoking
and lung cancer and based upon the results thus obtained from sample
studies, a valid statement about the association of smoking with lung cancer
in the whole population can be made.
(c) 30 per cent of all television viewers watched the show 20/20 last night. This
statement can be compared with the following statement: 30 per cent of those
who were interviewed watched the show 20/20 last night. The latter statement
is descriptive statistics since it is only presenting the data in a summarized
form. However, if we infer from the second statement to reach at the first
statement, then the first statement is an example of statistical inference.
Self-Instructional
188 Material
(d) Suppose that the Chancellor of Punjab University wanted to conduct a Data Analysis
Self-Instructional
190 Material
and rigid techniques of data collection. These are based on pre-formulated questions Data Analysis
and their responses. The outcomes of quantitative analysis are mostly broad based,
reliable and general in nature, and serves as the foundation for future action. The
limitations include its restrictive use due to statistics involved and struggles with
newer undiscovered phenomenon. NOTES
Qualitative data analysis refers to the process where descriptions are used
instead of numbers and values to interpret and present data. This type of data
analysis brings in to use rather flexible yet methodological methods of data collection
and interpretation. It seeks to identify the underlying reasons for a phenomenon
and the results obtained through this process is often used as hypothesis for
quantitative data analysis. The methods of data collection in qualitative data analysis
elicits unlimited and varied responses. There are many different methods of
collecting qualitative data including techniques like documents, observations,
interviews, etc. The outcomes of qualitative data analysis is mostly exploratory or
investigation and not conclusive in nature. The limitations of this analysis are that
the results cannot be applied to general population, it struggles with application of
statistics and its effectiveness due to the instruments used may suffer.
Parametric vs Non-Parametric Tests and Conditions for Satisfaction
Simply defined, it refers to the tests which makes assumptions about the parameters
of the population. Parametric tests are used for data with normal distribution. It
uses the measures of for ratio or interval data. The mean is mostly the central
measure. The information about the population is entirely known and specific
assumptions are made about the population. It is used in quantitative data analysis.
The parametric tests are applicable for variables only. These are considered more
powerful in comparison when the assumptions are met. Examples include paired
and unpaired t-tests, Pearson Correlation, Analysis of Variance, etc.
Non-parametric tests are used for data of any distribution. It uses measures
of ordinal or nominal data. The central measure is usually the median. There is no
information available about the population and this test is assumption free. It is
used for quantitative, qualitative and ranked data. The non-parametric tests apply
both to variables and attributes. It is easier to calculate. There are many examples
including Wilcoxon Rank Sum test, Mann-Whitney U test, Spearman Correlation,
Kruskal Wallis test among others.
Self-Instructional
Material 191
Data Analysis
9.4 ANSWERS TO CHECK YOUR PROGRESS
QUESTIONS
NOTES 1. One limitation of the nominal scale measurement is that we cannot say
whether the assigned number to an object is higher of lower than the one
assigned to another option. The ordinal scale measurement takes care of
this limitation.
2. The most common example of interval scale data is the relationship between
Celsius and Farenheit temperature.
3. Ration scale is the highest level of measurement and takes care of the
limitations of the interval scale measurement.
4. Averages and proportions are examples of descriptive statistics.
5. The measured characteristics of the sample are known as sample statistics,
while the measured characteristics of the population are known as population
parameters.
6. Examples of parametric tests include paired and unpaired t-tests, Pearson
Correlation, Analysis of Variance, etc.
9.5 SUMMARY
Self-Instructional
192 Material
An attitude is viewed as an enduring disposition to respond consistently in a Data Analysis
given manner to various aspects of the world, including persons, events and
objects. A company is able to sell its products or services when its customers
have a favourable attitude towards its products/services. In the reverse
scenario, the company will not be able to sustain itself for long. It, therefore, NOTES
becomes very important to measure the attitude of the customers towards
the company’s products/services.
As the name suggests, descriptive statistics merely describe the data and
consist of methods and techniques used in collection, organization,
presentation and analysis of data in order to describe the various features
and characteristics of such data. These methods can either be graphical or
computational. Thus data can be presented in the form of a chart or a table
in order to show certain trends, proportions, maximum and minimum values
and so on.
Inferential statistics can be defined as those methods that are used to estimate
a characteristic of a population or making of a decision concerning a
population on the basis of the results obtained from a sample taken from the
same population.
Quantitative data analysis is the interpretation of data which involves numbers
or where numerical data is analysed.
Qualitative data analysis refers to the process where descriptions are used
instead of numbers and values to interpret and present data.
Simply defined, it refers to the tests which makes assumptions about the
parameters of the population. Parametric tests are used for data with normal
distribution
Non-parametric tests are used for data of any distribution. It uses measures
of ordinal or nominal data.
Self-Instructional
Material 193
Data Analysis
9.7 SELF ASSESSMENT QUESTIONS AND
EXERCISES
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
194 Material
Qualitative Data Analysis
10.0 INTRODUCTION
You were introduced to the qualitative method of data analysis in the previous unit.
In comparison to quantitative data, qualitative data analysis makes use of words
and symbols instead of numbers to make an analysis. This aspect of making sense
of subjective rather worded data makes the process seem less stringent. But even
then, qualitative analysis does follow a discipline of making assumptions of large
set of data to make an empirical analysis. And even though, compared to quantitative
analysis, qualitative analysis looks less linear and more iterative. In this unit, you
will learn certain important aspects or tools crucial to qualitative analysis of data.
10.1 OBJECTIVES
Rural Urban
Population
Males Females
Self-Instructional
198 Material
limits is known as class magnitudes. Class limits may be generally be stated in any Qualitative Data Analysis
Self-Instructional
200 Material
tables also known as manifold tables which give information about several Qualitative Data Analysis
NOTES 1. In case of hand coding some standard method may be used. One such
standard method is to code in the margin with a coloured pencil. The other
method can be to transcribe the data from questionnaire to a coding sheet.
Whatever method is adopted one should see that coding errors are altogether
eliminated or reduced to minimum level.
2. When data are classified by the presence and absence of an attribute it is
known as simple classification.
3. In case of exclusive class intervals the upper limit of a class interval is the
lower limit of the succeeding class interval.
4. Constant comparison is normally associated with grounded theory.
10.4 SUMMARY
Self-Instructional
202 Material
of summarizing raw data and displaying the same in compact for further Qualitative Data Analysis
Self-Instructional
Material 203
Qualitative Data Analysis Long Answer Questions
1. Explain the objectives and essentials of classification of data.
2. Discuss in detail the methods of classification of data.
NOTES 3. Describe the essentials of tabulation.
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J. P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
204 Material
Analysis and Interpretation
INTERPRETATION
NOTES
OF DATA
Structure
11.0 Introduction
11.1 Objectives
11.2 Concept of Parameter and Statistics
11.2.1 Levels of Confidence
11.2.2 Degrees of Freedom
11.2.3 Standard Error of Mean
11.2.4 Two-Tailed and One-Tailed Tests of Significance
11.2.5 t-Test (Independent and Correlated Samples)
11.2.6 ANOVA: Assumptions
11.2.7 Correlations
11.3 Answers to Check Your Progress Questions
11.4 Summary
11.5 Key Words
11.6 Self Assessment Questions and Exercises
11.7 Further Readings
11.0 INTRODUCTION
Statistics has become an integral part of our daily lives. Every day, we are confronted
with some form of statistical information through newspapers, magazines and other
forms of communication. Such statistical information has become highly influential
in our lives. Indeed, the famous science fiction writer H.G. Wells had predicted
nearly a century ago that statistical thinking will one day be as necessary for efficient
citizenship as the ability to read and write. Thus, the subject of statistics in itself,
has gained considerable importance in affecting the processes of our thinking and
decision-making.
In this unit, you will learn about the concept of parameter and statistics
which includes the levels of confidence, degrees of freedom, standard error of
mean and the tests of significance.
11.1 OBJECTIVES
( – )
N
Self-Instructional
Material 207
Analysis and Interpretation Example 11.1: Suppose a babysitter has 5 children under her supervision with
of Data
average age of 6 years. However, individually, the age of each child be as follows:
X1 = 2
NOTES X2 = 4
X3 = 6
X4 = 8
X 5 = 10
Now these 5 children would constitute our entire population, so that N = 5.
Solution:
X
The population mean μ
N
2 4 6 8 10
= 30 / 5 6
5
and the standard deviation is given by the formula:
( μ) 2
σ
N
Total = ( X – )2 = 40
Then,
40
σ 8 2.83
5
Now, let us assume the sample size, n = 2, and take all the possible samples of
size 2, from this population. There are 10 such possible samples. These are as
follows, along with their means.
X1, X 2 (2, 4) X1 = 3
X1, X 3 (2, 6) X2 = 4
X1, X 4 (2, 8) X3 = 5
X1, X 5 (2,10) X4 = 6
Self-Instructional
208 Material
Analysis and Interpretation
X 2, X 3 (4, 6) X5 = 5 of Data
X 2, X 4 (4, 8) X6 = 6
X 2, X 5 (4, 10) X7 = 7
X 3, X 4 (6, 8) X8 = 7 NOTES
X 3, X 5 (6, 10) X9 = 8
X 4, X 5 (8, 10) X 10 = 9
Now, if only the first sample was taken, the average of the sample would be
3. Similarly, the average of the last sample would be 9. Both of these samples are
totally unrepresentative of the population. However, if a grand mean X of the
distribution of these sample means is taken, then,
10
X
i 1
i
X
10
34 56567 7 89
60 / 10 6
10
This grand mean has the same value as the mean of the population. Let us
organize this distribution of sample means into a frequency distribution and
probability distribution.
Sample Mean Freq. [Link]. Prob.
3 1 1/10 .1
4 1 1/10 .1
5 2 2/10 .2
6 2 2/10 .2
7 2 2/10 .2
8 1 1/10 .1
9 1 1/10 .1
1.00
Self-Instructional
Material 209
Analysis and Interpretation Sample mean ( X ) Prob. P( X )
of Data
3 .1
4 .1
NOTES 5 .2
6 .2
7 .2
8 .1
9 .1
1.00
Then,
XP( X ) = (3 × .l) + (4 × .l) + (5×.2) + (6 × .2) + (7 × .2) + (8 × .l)
+ 9 ×.l) = 6
This value is the same as the mean of the original population.
(ii) The spread of the sample means in the distribution is smaller than in the
population values. For example, the spread in the distribution of sample
means above is from 3 to 9, while the spread in the population was from
2 to 10.
(iii) The shape of the sampling distribution of the means tends to be, ‘Bell-
shaped’ and approximates the normal probability distribution, even when
the population is not normally distributed. This last property leads us to
the ‘Central Limit Theorem’.
Thus we can calculate –× for Example 11.1 of the sampling distribution of
the ages of 5 children as follows:
X (μ ) ( X μ )2
3 6 9
4 6 4
5 6 1
6 6 0
7 6 1
8 6 4
9 6 9
( X μ ) 2 = 28
Then,
( – )
N
28
7
42
Self-Instructional
210 Material
However, since it is not possible to take all possible samples from the Analysis and Interpretation
of Data
population, we must use alternate methods to compute –× .
The standard error of the mean can be computed from the following formula,
if the population is finite and we know the population mean. Hence,
NOTES
σ ( N n)
σ
n ( N 1)
Where,
= Population standard deviation
N = Population size
n = Sample size
This formula can be made simpler to use by the fact that we generally deal
with very large populations, which can be considered infinite, so that if the population
size A’ is very large and sample size n is small, as for example in the case of items
tested from assembly line operations, then,
( N – n)
would approach 1.
( N –1)
Hence,
n
( N – n)
The factor ( N – n)
is also known as the ‘finite correction factor’, and should be
used when the population size is finite.
As this formula suggests, –× decreases as the sample size (w) increases,
meaning that the general dispersion among the sample means decreases, meaning
further that any single sample mean will become closer to the population mean, as
the value of (–×) decreases. Additionally, since according to the property of the
normal curve, there is a 68.26 per cent chance of the population mean being
within one –× of the sample mean, a smaller value of –× will make this range
shorter; thus making the population mean closer to the sample mean (refer Example
11.2).
Example 11.2: The IQ scores of college students are normally distributed with
the mean of 120 and standard deviation of 10.
(a) What is the probability that the IQ score of any one student chosen at
random is between 120 and 125?
(b) If a random sample of 25 students is taken, what is the probability that the
mean of this sample will be between 120 and 125.
Self-Instructional
Material 211
Analysis and Interpretation Solution:
of Data
(a) Using the standardized normal distribution formula,
NOTES
125
= 120
= 10
( X – )
Z
125 –120
Z 5 / 10 .5
10
The area for Z = .5 is 19.15.
This means that there is a 19.15 per cent chance that a student picked up at
random will have an IQ score between 120 and 125.
11.2.4 Two-Tailed and One-Tailed Tests of Significance
Suppose a null hypothesis were set up that there was no difference other than a
sampling error difference between the mean height of two groups, A and B. We
would be concerned only with the difference and not with the superiority or inferiority
in height of either group. To test this hypothesis, we apply two-tailed test as the
difference between the obtained means of height of two groups may be as often in
one direction (plus) as in the other (minus) from the true difference of zero. Moreover,
for determining probability, we take both tails of sampling distribution.
For a large sample two-tailed test, we make use of a normal distribution
curve. The 5 per cent area of rejection is divided equally between the upper and
the lower tails of this curve and we have to go out to ±1.96 on the base line of the
curve to reach the area of rejection as shown in the Figure 11.1.
Self-Instructional
212 Material
Analysis and Interpretation
of Data
NOTES
Rejection Area Acceptance Area
2.5% 47.50% 47.50% 2.5%
- - - - - - - - - -- - -95%- - - - - - - - - - - - -
Fig. 11.1 A Two-Tailed Test at 0.05 Levels (2.5 Per Cent at Each Level)
Similarly, if we have 0.5 per cent area at each end of the normal curve
where 1 per cent area of rejection is to be divided equally between its upper and
lower tails, it is necessary to go out to ± 2.58 on the base line to reach the area of
rejection as shown in the Figure 11.2.
Acceptance Area
49.50% 49.50%
Fig. 11.2 A Two-Tailed Test at 0.01 level (0.5 Per Cent at Each Level)
In the case of the above example, a null hypothesis was set up where there
was no difference other than a sampling error difference between the mean creative
thinking [Link]. score of males and females. Thus, we were concerned, with a
difference and not in superiority or inferiority of either group in the creative thinking
ability. To test this hypothesis, we applied ‘two-tailed test’ as the difference between
the two means might have been in one direction (plus) or in the other (minus) from
the true difference of zero and we took both tails of sampling distribution in
determining probabilities.
Self-Instructional
Material 213
Analysis and Interpretation As is evident from the above example, we make use of a normal distribution
of Data
curve in the case of a large sample ‘two-tailed test’. The 5 per cent area of rejection
is equally divided between the upper and lower tails of the curve and we have to
go out to ±1.96 on the base line of the curve to reach the area of rejection.
NOTES
Similarly, we have 0.5 per cent area at each end of the normal curve when
1 per cent of rejection is to be divided equally between its upper and lower tails
and it is necessary to go out to ± 2.58 on the base line to reach the area of
rejection.
In the above problem, if we change the null hypothesis as: male group of
[Link]. have significantly higher creative thinking than that of the female group; or
male group have significantly lower creative thinking than the female group of
[Link]. course, then each of these hypotheses indicates a direction of difference. In
such situations, the use of ‘one-tailed test’ is made. For such a test, the 5 per cent
area or 1 per cent area of rejection is either at the upper tail or at the lower tail of
the curve, to be read from 0.10 column (instead of 0.05) and 0.02 column (instead
of 0.0l).
Application of t-Test for Testing the Significance of Difference
Between Two Independent Small Samples
The frequency distribution of small sample means drawn from the same population
forms a t-distribution, and it is reasonable to expect that the sampling distribution
of the difference between the means computed from two different populations will
also fall under the category of t-distribution. Fisher provided the formula for testing
the difference between the means computed from independent small samples as
follows.
M1 M 2
t
x12 x22 N N
1 2 (11.3)
N1 N 2 2 N1 N 2
Where,
M1 and M2 = Means of two samples.
2 2
x1 and x2 = Sums of squares of the deviations from the means in the
two samples.
N1 and N2 = Number of cases in the two samples.
df = Degrees of freedom = N1 + N2 – 2.
To illustrate the use of the Formula (5), let us test the significance of the
difference between mean scores of 7 boys and 10 girls in an intelligence test as
illustrated in Table 11.1.
Self-Instructional
214 Material
Table 11.1 Scores of 7 Boys and 10 Girls in an Intelligence Test Analysis and Interpretation
of Data
(N1 = 7) (N2 = 10)
2
Boys x1 x1 Girls x2 x22
X1 X2 NOTES
13 0 0 10 –4 16
14 1 1 16 2 4
11 –2 4 12 –2 4
12 –1 1 13 –1 1
15 2 4 18 4 16
13 0 0 13 –1 1
13 0 0 19 5 25
14 0 0
13 –1 1
12 –2 4
X1 = 91 X12 = 10 X2 = 140 X22 = 72
91
M1 13
7
140
M2 14
10
df = N1 + N2 – 2 = 7 + 10 – 2 = 15
Using Formula (5) we compute t as follows:
14 13
t
10 72 7 10
7 10 2 7 10
1
82 17
15 10
1
9.29
0.33
To test the significance of difference between the two means by making use
of the ‘two-tailed test’ (null hypothesis, i.e., no differences between the two groups),
we look for the t-critical values for rejection of null hypothesis for (7 + 10 – 2) or
15 df. These t-values are 2.13 at 0.05 and 2.95 at 0.01 levels of significance.
Since the obtained t-value 0.33 is less than the table value necessary for the rejection
of the null hypothesis at 0.05 levels for ‘df’ 15, the null hypothesis is accepted and
it may be concluded that there is no significant difference in the mean intelligence
scores of males and females. If we change the null hypothesis as: boys will have
Self-Instructional
Material 215
Analysis and Interpretation higher intelligence scores than girls or males will have lower intelligence scores
of Data
than females, then each of these hypotheses indicates a direction of difference
rather than simply the existence of the difference. So, we make use of ‘one-tailed
test’. For given degrees of freedom, i.e., 15, the 0.05 level is read from the 0.10
NOTES column ( p/2 = 0.05) and the 0.01 level from 0.02 column( p/2 = 0.01) of the ‘t’
table. In the one-tailed test, for 15 ‘df’ t-critical values at 0.05 and 0.01 levels, as
read from the 0.10 and the 0.02 columns are 1.75 and 2.60, respectively. Since
the computed t-value of 0.33 does not reach the table value at 0.05 levels (i.e.,
1.75 for 0.l0), we may conclude that the difference in two groups is present merely
because of chance factors.
11.2.5 t-Test (Independent and Correlated Samples)
Sir William S. Gosset (pen name Student) developed a significance test and through
it made a significant contribution to the theory of sampling applicable in case of
small samples. When population variance is not known, the test is commonly known
as Student’s t-test and is based on the t distribution.
Like normal distribution, t distribution is also symmetrical but happens to be
flatter than normal distribution. Moreover, there is a different t distribution for
every possible sample size. As the sample size gets larger, the shape of the t
distribution loses its flatness and becomes approximately equal to the normal
distribution. In fact, for sample sizes of more than 30, the t distribution is so close
to the normal distribution that we will use the normal to approximate the t distribution.
Thus, when n is small, the t distribution is far from normal, but when n is infinite, it
is identical to normal distribution.
For applying t-test in context of small samples, the t value is calculated
first of all and then the calculated value is compared with the table value of t
at certain level of significance for given degrees of freedom. If the calculated
value of t exceeds the table value (say t0.05), we infer that the difference is
significant at 5 per cent level but if calculated value is t0 is less than its concerning
table value, the difference is not treated as significant.
The t-test is used when two conditions are fullfiled,
(a) The sample size is less than 30, i.e., when n 30..
(b) The population standard deviation (p) must be unknown.
In using the t-test, we assume the following:
(a) That the population is normal or approximately normal.
(b) That the observations are independent and the samples are randomly drawn
samples.
(c) That there is no measurement error.
(d) That in the case of two samples, population variances are regarded as equal
if equality of the two population means is to be tested.
Self-Instructional
216 Material
The following formulae are commonly used to calculate the t value: Analysis and Interpretation
of Data
(a) To Test the Significance of the Mean of a Random Sample
| X |
t NOTES
S | SEx X
( X i X )2
n
SEX s
n n
and the degrees of freedom = (n – 1).
The above stated formula for t can as well be stated as under:
| X | | X | | X |
t = n
SE X ( X X ) 2
( X X )2
n 1 n 1
n
If we want to work out the probable or fiducial limits of population mean
(µ) in case of small samples, we can use either of the following:
(a) Probable limits with 95 per cent confidence level:
X SE X (t0.05 )
At other confidence levels, the limits can be worked out in a similar manner,
taking the concerning table value of t just as we have taken t0.05 in (a) and t0.01 in
(b) above.
(b) To Test the Difference between the Means of the Two Samples
| X1 X 2 |
t
SE X 1 X 2
Example 11.4: Two random samples have been selected and a two sample
t-test for the difference in population means is conducted with H0: 1 = 2 vs. Ha:
1 > 2. The results are s1 = 2.5, x1 = 9.7, n1 = 30, s2 = 2.9, x2 = 9.1 and n2 = 35.
What is the value of t-test?
Solution: The value of t-test is obtained as follows:
Here,
x1 = 9.7
x2 = 9.1
n1 = 30
Self-Instructional
Material 219
Analysis and Interpretation n2 = 35
of Data
s1 = 2.5
s2 = 2.9
NOTES We know:
x1 x2
t
s12 s22
n1 n 2
= 9.7 – 9.1 / (2.5)2/30 + (2.9)2 / 35
= 0.6 / 6.25/ 30 + 8.41/35
= 0.6 0.208 + 0.240
= 0.6 / 0.448
= 0.6 / 0.66
= 0.90
The value of t as per as per t-test statistics is 0.90.
t-Test for Independent and Dependent Group
Under this heading, you will learn about t-Test for independent and dependent
group.
Independent t-Test
The sampling distribution of the difference between the means of two independent
samples provides the basics for the testing of a mean difference hypothesis between
two groups.
A typical research in which one would use the independent t -test might
involve one group of employees receiving sales training and the second group of
employees not receiving any sales training. The number of sales for each group is
recorded and averaged. The null hypothesis would be stated when the average
sales for the two groups are equal. The alternative hypothesis would be stated
when the group receiving the sales training will on average have higher sales than
the group that did not receive any sales training. If the sample data for the two
groups were recorded for Sales Training where the mean = 40, standard deviation=
10, n = 100, then the independent t-test can be computed as follows:
X1 X 2
t
S X1 X 2
D
t
SD
The numerator in the above formula is the average difference between the post
and pre means score on the attribute towards violence inventory which is 73 –
67.5 = 5.5.
The denominator is calculated as follows:
Student Pre Post D D2
Self-Instructional
Material 221
Analysis and Interpretation 11.2.6 ANOVA: Assumptions
of Data
In business decisions, we are often involved in determining if there are significant
differences among various sample means, from which conclusions can be drawn
NOTES about the differences among various population means. For example, we may be
interested to find out if there are any significant differences in the average sales
figures of 4 different salesman employed by the same company, or we may be
interested to find out if the average monthly expenditures of a family of 4 in 5
different localities are similar or not, or the telephone company may be interested
in checking, whether there are any significant differences in the average number of
requests for information received in a given day among the 5 areas of City (Under
Study), and so on. The methodology used for such types of determinations is
known as ANalysis Of VAriance or ANOVA. This technique is one of the most
powerful techniques in statistical analysis and was developed by R.A. Fisher. It is
also called the F-Test.
There are two types of classifications involved in the analysis of variance.
The one-way analysis of variance refers to the situations when only one fact or
variable is considered. For example, in testing for differences in sales for three
salesman, we are considering only one factor, which is the salesman’s selling ability.
In the second type of classification, the response variable of interest may be affected
by more than one factor. For example, the sales may be affected not only by the
salesman’s selling ability, but also by the price charged or the extent of advertising
in a given area.
The Basic Principle of ANOVA
The basic principle of ANOVA is to test for differences among the means of the
populations by examining the amount of variation within each of these samples,
relative to the amount of variation between the samples. In terms of variation
within the given population it is assumed that the values of (Xij) differ from the
mean of this population only because of random effects i.e., there are influences
on (Xij) which are unexplainable, whereas in examining differences between
populations we assume that the difference between the mean of the jth population
and the grand mean is attributable to what is called a ‘specific factor’ or what is
technically described as treatment effect. Thus, while using ANOVA, we assume
that each of the samples is drawn from a normal population and that each of these
populations has the same variance. We also assume that all factors other than the
one or more being tested are effectively controlled. This, in other words, means
that we assume the absence of many factors that might affect our conclusions
concerning the factor(s) to be studied.
Thus, a composite procedure for testing simultaneously the difference
between several sample means is known as the ANOVA. It helps us to know
whether any of the differences between the means of the given sample are significant.
If the answer is yes, we examine pairs (with the help of the t-test) to see just where
Self-Instructional
the significant difference lie. If the answer is no, we do not proceed further.
222 Material
In such a test, as the name implies, we usually deal with the analysis of the Analysis and Interpretation
of Data
variances. Variances are simply the arithmetic average of the squared deviation
from their means. In other words, it is the square of standard deviation (Variance
= Variance has a quality which makes it especially useful. It has an additive
property, which the standard deviation with its square root does not possess. NOTES
Variance on this account can be added up and broken down into components.
Hence, the term ‘analysis of variance’ deals with the task of analyzing of breaking
up the total variance of a large sample or a population consisting of a number of
equal groups or sub-samples into two components (two kinds of variances), given
as follows:
‘Within Groups’ Variance: This is the average variance of the members
of each group around their respective group means, i.e., the mean value of
the scores in a sample (as members of each group may vary among
themselves).
‘Between Groups’ Variance: This represents the variance of group means
around the total or grand mean of all groups, i.e., the best estimate of the
population mean (as the group means may vary considerably from each
other).
The technique of analysis of variance is applied to determine if any two of the
seven means differ significantly from each other by a single test, known as F-test,
rather than 21 t-tests. The F-test makes it possible to determine whether the sample
means differ from one another (between group variance) to a greater extent than the
test scores differ from their own sample means (within group variance) using the
ratio given below:
Variance between the groups
F
Variance within groups
Self-Instructional
Material 223
Analysis and Interpretation One-Way ANOVA
of Data
One-way analysis of variance, also abbreviated as one-way ANOVA, is a
technique used to compare means of two or more samples using the F distribution.
NOTES This technique can be used only for numerical data.
Example 11.5: To illustrate the use of F-test, let us consider an example of 20
students who have been randomly assigned to four groups of five each, to be
taught by different methods, i.e., A, B, C and D. Their performance scores on an
achievement test, administered after the completion of experiment are given in
Table 11.2.
Table 11.2 Achievement Test Scores of the Four Groups Taught
through Four Different Methods
Methods or Groups
A B C D
(X1) (X2) (X3) (X4)
14 19 12 17
15 20 16 17
11 19 16 14
10 16 15 12
12 16 12 17
X 62 90 71 77 300
2
X 786 1634 1025 1207 4652
Using formula,
27.60
F 6.374
4.33
In the present problem, the null hypothesis asserts that four sets of scores
are in reality the scores of four random samples drawn from the same normally
distributed schools, and that the means of the four groups A, B, C, and D will
differ only through fluctuations of sampling. For testing this hypothesis, we divided
the ‘between means’ variance by the ‘within treatments’ variance and compared
the resulting variance ratio, called F, with the F-values. The F value of 6.374 in
the present case is to be checked for table value for ‘df’ 3 and 16 (the degrees of
freedom for numerator and denominator). The table values for 0.05 and 0.01
levels of significance are 3.24 and 5.29. Since the computed F-value of 6.374 is
greater than the table values, we reject the null hypothesis and conclude that the
means of the four groups differ significantly.
Self-Instructional
Material 225
Analysis and Interpretation Two-Way ANOVA
of Data
We have studied one way ANOVA involving four different methods of teaching.
In two-way analysis of variance classification, an estimate of population variance,
NOTES i.e., total variance is supposed to be broken up into (i) Variance due to adjustment,
(ii) Variance due to anxiety alone, and (iii) The residual variance called interaction
variance (Adj × Anx), where A = Adjustment and Anx = Anxiety.
Example 11.6: A study has been conducted on anxiety and adjustment with the
help of 2 × 2 factorial designs. It has four conditions and the score is given below
in the Table.
Table Score
Adjustment
High Low
High A B
Anxiety
Low C D
N1 = 10 N2 = 10 N3 = 10 N4 = 10
Self-Instructional
226 Material
2
Analysis and Interpretation
X 1 X 2 X 3 X 4 of Data
C
N
2
58 42 63 56 47961
C NOTES
40 40
1199.02
Total SS = X1 2 + X2 2 + X3 2 + X4 2 + …. – Correction
= 338 + 188 + 405 + 318 – 1199.02
= 1249 – 1199.02 = 49.98
2 2 2 2
X 1 X 2 X 3 X 4
Among SS = Correction
N1 N2 N3 N4
2 2 2 2
58 42 63 56
= Correction
10 10 10 10
= 3364 + 1764 + 3969 + 3136/40 – Correction
= 336.4 + 176.4 + 396.9 + 313.6 – Correction
= 1223.3 – 1199.02 = 24.28
Within SS = Total SS – Among SS
= 49.98 – 24.28 = 25.70
SS between amount of first IV (Adjustment)
2 2
X 1 X 3 X 2 X 4
= Correction
N1 N3 N2 N4
2 2
58 63 42 56
= Correction
10 10 10 10
= 1212.25 1199.02
= 13.23
SS between amount of second IV (Anxiety)
2 2
X 1 X 2 X 3 X 4
= Correction
N1 N 2 N3 N4
2 2
58 42 63 56
= 1199.0211
10 10 10 10
10000 14161
= 1199.0211
20 20
= 500 + 708.05 – 1192.02
= 1208.05 – 1192.02
= 9.03 Self-Instructional
Material 227
Analysis and Interpretation Interaction SS = Among SS – Between SS for first IV – Between second
of Data
IV = 24.28 – 13.23 – 9.03 = 2.02.
Summary: Analysis of Variance
Self-Instructional
Material 229
Analysis and Interpretation 2. Negative correlation: When two variables X and Y move in the opposite
of Data
direction, the correlation is negative. If one variable increases, the other decreases
and vice versa. The examples of negative correlation are usually the quantity
demanded and the price of the commodity. The scatter of the points on the variables
NOTES X and Y is clustered around a negatively sloped straight line/curve in such a situation
as shown in Figure 11.4. In the figure, we find that the variables X and Y are
moving in the opposite direction.
3. Zero correlation: The correlation between two variables X and Y is zero
when the variables move in no connection with each other. If the variable X increases,
Y may increase or decrease in some situation. The scatter of the points of the
variables X and Y in case of zero correlation is given in Figure 11.5. Zero correlation
does not mean that the variables are not related. We are, here, dealing with a
linear correlation and there could be a non-linear relation between them.
Quantitative Estimate of a Linear Correlation
A quantitative estimate of a linear correlation between two variables X and Y is
given by Karl Pearson as:
(11.4)
(11.5)
It may be noted that the above-mentioned formulae are for the linear
correlation coefficient. The linear correlation coefficient takes a value between –1
and +1 (both values inclusive). If the value of the correlation coefficient is equal to
1, the two variables are perfectly positively correlated and the scatter of the points
of the variables X and Y will lie on a positively sloped straight line. Similarly, if the
correlation coefficient between the two variables X and Y is –1, the scatter of the
points of these variables will lie on a negatively sloped straight line and such a
correlation will be called a perfectly negative correlation. It may be noted that the
Self-Instructional
230 Material
closer the scatter of points to the line, higher is the degree of correlation between Analysis and Interpretation
of Data
the variables.
Testing the Significance of the Correlation Coefficient
The statistical test for the significance of a correlation coefficient is conducted NOTES
using a t-statistic. The hypothesis to be tested is mentioned below:
(11.6)
Self-Instructional
Material 231
Analysis and Interpretation
of Data 11.4 SUMMARY
Self-Instructional
232 Material
Analysis and Interpretation
11.5 KEY WORDS of Data
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
Material 233
Statistical Analysis
BLOCK - IV
STATISTICAL ANALYSIS, RESEARCH REPORT AND
COMPUTER IN EDUCATIONAL RESEARCH
NOTES
12.0 INTRODUCTION
12.1 OBJECTIVES
Self-Instructional
234 Material
Statistical Analysis
12.2 NON-PARAMETRIC STATISTICS AMD SIMPLE
APPLICATIONS
Non-parametric tests are not based on set parameters or assumptions for testing. NOTES
Advantages and Disadvantages of Non-Parametric Tests
There are many advantages of a non-parametric test. These are:
They can be applied to many situations as they do not have the rigid
requirements of their parametric counterparts, like the sample having been
drawn from the population following a normal distribution. A researcher
can encounter an application where a numeric observation is difficult to
obtain but a rank value is not. For example, it is easy to obtain the rank data
on the preference of consumer for the various brands of toothpaste rather
than assigning a numerical value to them. By using ranks, it is possible to
relax the assumptions regarding the underlying populations.
Non-parametric tests can often be applied to the nominal and ordinal data
that lack exact or comparable numerical values. For example, the
respondents may be asked a question on their religion—Hindu, Sikh,
Christian, or Muslim. This is a nominal scale data and can only be analysed
by non-parametric methods.
Non-parametric tests involve very simple computations compared to the
corresponding parametric tests.
However, the methods are not without their own drawbacks and there are certain
disadvantages of non-parametric tests. These are:
A lot of information is wasted because the exact numerical data is reduced
to a qualitative form. For example, in one of the non-parametric tests like
the sign test, the increase or the gain is denoted by a plus sign whereas a
decrease or loss is denoted by a negative sign. No consideration is given to
the quantity of the gain or loss. A gain of `1 or `1 lakh would both receive
a plus sign.
Non-parametric methods are less powerful than parametric tests when the
basic assumptions of parametric tests are valid. Therefore, there is more
risk of accepting a false hypothesis and thus committing a type II error.
Null hypothesis in a non-parametric test is loosely defined as compared to
the parametric tests. Therefore, whenever the null hypothesis is rejected, a
non-parametric test yields a less precise conclusion as compared to the
parametric test. For example, corresponding to the null hypothesis that the
means of the two populations are equal in the parametric test, the null
hypothesis in a non-parametric test is that the two populations have same
probability distributions.
Self-Instructional
Material 235
Statistical Analysis In such a situation, rejecting a null hypothesis under the parametric test
would imply that the means of the two populations are different whereas
under a non- parametric test, it means that the two population distributions
are different but the specific form of the difference between the two
NOTES populations is not clearly defined.
In this section, we will discuss non-parametric tests such as chi-square, run
test, sign test, the Mann-Whitney U test, the Wilcoxon matched-pair rank
test and the Kruskal–Wallis test. The differences between parametric and
non-parametric tests are summarized below.
NOTES
Self-Instructional
Material 237
Statistical Analysis Make a note of the observed counts of the data points falling in different
cells
Compute the chi-square value given by the formula.
NOTES
where,
Oi = Observed frequency of ith cell
Ei = Expected frequency of ith cell
k = Total number of cells
k–1 = Degrees of freedom
Compare the sample value of the statistic as obtained in previous step with
the critical value at a given level of significance and make the decision.
A goodness of fit test is a statistical test of how well the observed data
supports the assumption about the distribution of a population. The test also
examines that how well an assumed distribution fits the data. Many a times, the
researcher assumes that the sample is drawn from a normal or any other distribution
of interest. A test of how normal or any other distribution fits a given data may be
of some interest.
Consider for example the case of the multinomial experiment which is the
extension of a binomial experiment. In the multinomial experiment, the number of
the categories k is greater than 2. Further, a data point can fall into one of the k
categories and the probability of the data point falling in the ith category is a constant
and is denoted by pi where i = 1, 2, 3, 4, ..., k. In summary, a multinomial experiment
has the following features:
There are fixed number of trials.
The trials are statistically independent.
All the possible outcomes of a trial get classified into one of the several
categories.
The probbilities for the different categories remain constant for each trial.
Consider as an example that a respondent can fall into any one of the four
non-overlapping income categories. Let the probabilities that the respondent will
fall into any of the four groups may be denoted by the four parameters p1, p2, p3,
and p4. Given these, the multinomial distribution with these parameters, and n the
number of people in a random sample, specifies the probabilities of any combination
of the cell counts.
Given such a situation, we may use a multinomial distribution to test how
well the data fits the assumption of k probability p1, p2, ..., pk of falling into the k
cells. The hypothesis to be tested is:
Self-Instructional
238 Material
H0 : Probabilities of the occurrence of events E1, E2, ..., Ek are given by the Statistical Analysis
Solution:
Let
pv : Proportion of customers preferring vanilla flavour.
pc : Proportion of customers preferring chocolate flavour.
ps : proportion of customers preferring strawberry flavour.
pm : proportion of customers preferring mango flavour.
H0 : pv = 0.62, pc = 0.18, ps = 0.12, pm = 0.08
H1 : Proportions are not that specified in the null hypothesis
The expected frequencies corresponding to the various flavors under the
assumption that the null hypothesis is true are:
Vanilla = 200 × 0.62 = 124
Chocolate = 200 × 0.18 = 36
Strawberry = 200 × 0.12 = 24
Mango = 200 × 0.08 = 16
Self-Instructional
Material 239
Statistical Analysis The computed value of chi-square is 4.323.
Table (5 per cent) = 9.488 (see Annexure 3 at the end of the book.)
NOTES
Self-Instructional
240 Material
Solution: Statistical Analysis
Let
p1 = Proportion of fatalities on Monday
p2 = Proportion of fatalities on Tuesday NOTES
p3 = Proportion of fatalities on Wednesday
p4 = Proportion of fatalities on Thursday
p5 = Proportion of fatalities on Friday
p6 = Proportion of fatalities on Saturday
p7 = Proportion of fatalities on Sunday
H0 : p1 = p2 = p3 = p4 = p5 = p6 = p7 =
Self-Instructional
Material 241
Statistical Analysis
NOTES
Since the sample chi-square value is less than the tabulated 2, there is not
enough evidence to reject the null hypothesis as shown in the figure below.
The problem can also be worked out using the p-value approach. The
sample value of 2 = 9.233 with 6 df is less than the critical value 10.645, which
corresponds to an area of 10 per cent. Therefore, the p value in this problem is
greater than 10 per cent, which is higher than the level of significance = 0.05.
Therefore, the null hypothesis is accepted. This means that the accidents occur on
different days with equal frequencies.
A chi-square test for independence of variables
The chi-square test can be used to test the independence of two variables each
having at least two categories. The test makes a use of contingency tables also
referred to as cross-tabs with the cells corresponding to a cross classification of
attributes or events. A contingency table with 3 rows and 4 columns (as an example)
is shown.
Assuming that there are r rows and c columns, the count in the cell
corresponding to the ith row and the jth column is denoted by Oij, where i = 1, 2,
..., r and j = 1, 2, ..., c. The total for row i is denoted by Ri whereas that
corresponding to column j is denoted by Cj. The total sample size is given by n,
which is also the sum of all the r row totals or the sum of all the c column totals.
Self-Instructional
242 Material
The hypothesis test for independence is: Statistical Analysis
The degrees of freedom for the chi-square statistic are given by (r – 1) (c – 1).
For a given level of significance a, the sample value of the chi-square is
compared with the critical value for the degree of freedom (r – 1) (c – 1) to make
a decision.
The expected frequency in the cell corresponding to the ith row and the jth
column is given by:
where,
Ri = Total for the ith row,
Cj = Total for the jth column,
n= Total sample size.
Let us consider a few examples:
Example 12.3: A sample of 870 trainees was subjected to different types of
training classified as intensive, good and average and their performance was noted
as above average, average and poor. The resulting data is presented in the table
below. Use a 5 per cent level of significance to examine whether there is any
relationship between the type of training and performance.
Solution:
H0 : Attribute performance and the training are independent.
H1 : Attribute performance and the training are not independent.
Self-Instructional
Material 243
Statistical Analysis The expected frequencies corresponding the ith row and the jth column in the
contingency table are denoted by Eij, where i = 1, 2, 3 and j = 1, 2, 3.
NOTES
Self-Instructional
244 Material
The critical value of the chi-square at 5 per cent level of significance with 4 Statistical Analysis
degrees of freedom is given by 9.49. The sample value of the chi-square falls in
the rejection region as shown in the figure on next page.
Therefore, the null hypothesis is rejected and one can conclude that there is
NOTES
an association between the type of training and performance.
Using a p value approach, it can be seen that the computed value of chi-
square (107.39) with 4 df is higher than the critical value (13.28) at 1 per cent
level of significance. Therefore, the p value of this problem is less than 0.01 which
is far below the level of significance. Therefore, the null hypothesis is rejected.
This means that there is a relationship between the type of training and the
performance.
A chi-square test for the equality of more than two population proportions
In certain situations, the researchers may be interested to test whether the
proportion of a particular characteristic is the same in several populations. The
interest may lie in finding out whether the proportion of people liking a movie is the
same for the three age groups, 25 and under, over 25 and under 50, and 50 and
over. To take another example, the interest may be in determining whether in an
organization, the proportion of the satisfied employees in four categories—class I,
class II, class III, and class IV employees—is the same. In a sense, the question
of whether the proportions are equal is a question of whether the three age
populations of different categories are homogeneous with respect to the
characteristics being studied. Therefore, the tests for equality of proportions across
several populations are also called tests of homogeneity.
The analysis is carried out exactly in the same way as was done for the
other two cases. The formula for a chi-square analysis remains the same. However,
two important assumptions here are different.
(i) We identify our population (e.g., age groups or various class employees)
and the sample directly from these populations.
(ii) As we identify the populations of interest and the sample from them directly,
the sizes of the sample from different populations of interest are fixed. This
Self-Instructional
Material 245
Statistical Analysis is also called a chi-square analysis with fixed marginal totals. The hypothesis
to be tested is as under:
H0 : The proportion of people satisfying a particular characteristic is
the same in population.
NOTES
H1 : The proportion of people satisfying a particular characteristic is
not the same in all populations.
The expected frequency for each cell could also be obtained by using the
formula as explained earlier. There is an alternative way of computing the same,
which would give identical results. This is shown in the following example:
Example 12.4: An accountant wants to test the hypothesis that the proportion of
incorrect transactions at four client accounts is about the same. A random sample
of 80 transactions of one client reveals that 21 are incorrect; for the second client,
the number is 25 out of 100; for the third client, the number is 30 out of 90
sampled and for the fourth, 40 are incorrect out of a sample of 110. Conduct the
test at = 0.05.
Solution:
Let p1 = Proportion of incorrect transaction for 1st client
p2 = Proportion of incorrect transaction for 2nd client
p3 = Proportion of incorrect transaction for 3rd client
p4 = Proportion of incorrect transaction for 4th client
Let H0 : p1 = p2 = p3 = p4
H1 : All proportions are not the same.
The observed data in the problem can be rewritten as:
Self-Instructional
246 Material
In fact, the sum of each row/column in both the observed and expected Statistical Analysis
frequency tables should be the same. Here, a bit of discrepancy is found because
of the rounding of the error. It can be easily verified that the expected frequencies
in each cell would be the same using the formula as already explained. NOTES
Now the value of the chi-square statistic can be calculated as:
Therefore,
We need to know the lower and upper limit of the contingency coefficient
(C) to determine how strong is the relationship between age and preference. The
lower limit of C equals zero when 2 is zero. The 2 will take a value of zero when
the variables are independent. The upper limit of C when the number of rows is
equal to the number of columns is given by the expression:
Self-Instructional
Material 247
Statistical Analysis between 0 and 0.707. This means that there is a moderate relationship between
the variables.
Phi coefficient (): There is another statistic called the phi-coefficient which can
be used to determine the strength of a relationship only in a 2 × 2 contingency
NOTES
table. The phi-coefficient like the correlation coefficient can assume any value
between –1 and 1.
Phi-coefficient (Æ) may be computed by using the following formula:
Cramer’s V statistic: When the number of rows is not equal to the number of
columns, we may use the statistic called Cramer’s V statistic given by:
of runs (r) in the above sample of the 45 entrants of a restaurant is shown below:
MMFMFFFMMMMFFFMMFFFMMMMM FFMMMF
FFMFFFFFMMFFFFF
NOTES
The total number of runs is 16 as shown by the lines below the identical
symbols. In the above example:
n (Total size of the sample) = 45
n1 (Number of males in the sample) = 20
n2 (Number of females in the samples) = 25
r (Number of runs) = 16
Too many or too few runs in a sequence indicates a lack of randomness.
For large samples, either n1 > 20 or n2 > 20, the distribution of runs (r) is normally
distributed with mean:
mr =
Self-Instructional
Material 249
Statistical Analysis Assuming a 5 per cent level of significance, the critical value of Z is given by
± 1.96. As the absolute Z value is greater than the absolute critical value of Z, the
null hypothesis is rejected. Therefore, the sequence of this observation is not
randomly generated.
NOTES
The example discussed above clearly fits into two categories (nominal
measurement). The test for randomness can also be applied to the interval or ratio
scale data. What is required is that the interval/ratio scale data should be converted
into a nominal scale measurement. To partition the data into two categories, one
could use the value of mean or median and randomness can be tested for the
numerical data above or below the median. For illustration purposes, consider the
following example.
Example 12.6: The data listed below is the lifetime of batteries in hours produced
by ZIDA company in a particular order.
270, 280, 248, 260, 220, 285, 270, 266, 269, 266, 272
225, 228, 290, 284, 282, 276, 269, 250, 249, 262, 273
277, 258, 264, 269, 276, 278, 249, 286, 282, 264, 201
215, 222, 238, 212, 242, 236, 247, 249, 248, 256, 271
282, 305, 217, 303, 305, 309, 320, 262, 244, 262, 267
Assuming a significance level of 5 per cent, determine whether the sample
lifetime of the batteries produced by ZIDA is random.
Solution:
H0 : Lifetime of batteries is random.
H1 : Lifetime of batteries is not random.
There are 55 observations. We will first compute the median of the
distribution by arranging the data in an ascending order of magnitude shown below:
201, 212, 215, 217, 220, 222, 225, 228, 236, 238, 242
244, 247, 248, 248, 249, 249, 249, 250, 256, 258, 260
262, 262, 262, 264, 264, 266, 266, 267, 269, 269, 269
270, 270, 271, 272, 273, 276, 276, 277, 278, 280, 282
282, 282, 284, 285, 286, 290, 303, 305, 305, 309, 320
As there are 55 observations, the value of the middle (28th) observation
when data is arranged in an ascending order of magnitude gives the median of
distribution. Please note that the 28th observation when the data is arranged in an
ascending order of magnitude is 266. There are two observations having a value
of 266. Therefore, these two are discarded and for further analysis we will have
53 observations. Now the original data will be divided into two categories—
above the median denoted by (A) and below the median denoted by (B). The
number of runs could be obtained as shown below:
Self-Instructional
250 Material
AA B B B AAA B B AAAAA B B B AA B B AAA B AA Statistical Analysis
B B B B B B B B B B B B AAA B AAAA B B B A
The total number of runs (r) = 17
Number of observations above median (n1) = 26 NOTES
Number of observations below median (n2) = 27
Total number of observations (n) = 53
As both n1 and n2 are greater than 20, the distribution of runs (r) could be
approximated by normal distribution with mean:
The critical value of Z at 5 per cent level of significance equals 1.645. As the
sample value of Z is greater than the critical value, the null hypothesis is rejected
and the median of the distribution is greater than 19.
Example 12.8: A survey was conducted to understand the preference for fast
food by the inhabitants of a small town. A sample of 100 respondents indicated
that 54 do not prefer fast food whereas 46 have a preference for the fast food. By
using a sign test, examine the hypothesis that half of the inhabitants of the town
prefer fast food. Let the level of significance be 5 per cent.
Solution:
H0 : p = ½
H1 : p ½
where, p = Proportion not preferring fast food.
Denote those not preferring fast food by a plus sign and those preferring
fast food by a minus sign. Therefore, there are 54 plus signs and 46 minus signs.
The test statistic in this case is:
Solution:
H0 :There is no significant difference between the two versions.
H1 :There is a significant difference between the two versions.
We note that there are 7 plus signs (score of Version 1 is more than that of
Version 2), 9 minus signs (score of Version 1 is less than that of Version 2). There
is one case with an identical score and therefore, this observation is dropped from
the analysis and accordingly the sample size is reduced to 16.
Now, the Z statistic may be applied to test the hypothesis. This is because
both np and nq are greater than 5 (16 × ½ = 8);
Self-Instructional
254 Material
The critical value of Z at a 5 per cent level of significance is ± 1.96 (two- Statistical Analysis
tailed test). As the absolute sample value of Z is less than the absolute critical
value, there is not enough evidence to reject H0. Therefore, there is no statistical
difference between the IQ scores of the two versions. Therefore, it is safe to use
any of the versions for measuring IQ. NOTES
(iii) Define
and
Self-Instructional
Material 255
Statistical Analysis Please note that the following expression will hold true:
U1 + U2 = n1n2
Mann-Whitney test for a large sample: If n1 or n2 is greater than 10, a
NOTES large sample approximation can be used for the distribution of the Mann-Whitney
U statistic. For this purpose, either of U1 or U2 could be used for testing a one-
tailed or a two-tailed test. In this test, U2 will be used for the purpose.
Under the assumption that the null hypothesis is true, the U2 statistic follows
an approximately normal distribution with mean:
Solution:
H0 : Two populations have identical probability distributions.
H1 : Population A is shifted to the right of population B.
Self-Instructional
256 Material
We pool both the samples and rank them. This is shown below: Statistical Analysis
NOTES
Self-Instructional
Material 257
Statistical Analysis The critical value of Z at a 5 per cent level of significance is given by 1.645.
The sample value of Z exceeds the critical value of Z and the null hypothesis is
rejected. Therefore, Bank A has a larger number of bounced cheques as compared
to Bank B.
NOTES
12.2.5 Wilcoxon Signed-Rank Test for Paired Samples
The Mann-Whitney U test just discussed assumes that the two samples are
independent. However, there are instances when the sample data consists of paired
observations. Examples of paired samples include a study where husband and
wife are matched or where subjects are studied before and after experimentation
or observations are taken on a variable for brother and sister. The case of paired
sample (dependent sample) was discussed in Unit 11 using a t distribution. The
use of t distribution is based on the normality assumption. However, there are
instances when the normality assumption is not satisfied and one has to resort to a
non-parametric test. One such test earlier discussed was the two-sample sign test.
In this test, only the sign of the difference (positive or negative) was taken into
account and no weightage was assigned to the magnitude of the difference. The
Wilcoxon matched-pair signed rank test takes care of this limitation and attaches
a greater weightage to the matched pair with a larger difference. The test, therefore,
incorporates and makes use of more information than the sign test. This is, therefore,
a more powerful test than the sign test.
The test procedure is outlined in the following steps:
(i) Let di denote the difference in the score for the ith matched pair. Retain
signs, but discard any pair for which d = 0.
(ii) Ignoring the signs of difference, rank all the di’s from the lowest to highest.
In case the differences have the same numerical values, assign to them the
mean of the ranks involved in the tie.
(iii) To each rank, prefix the sign of the difference.
(iv) Compute the sum of the absolute value of the negative and the positive
ranks to be denoted as T– and T+ respectively.
(v) Let T be the smaller of the two sums found in step iv.
When the number of the pairs of observation (n) for which the difference is not
zero is greater than 15, the T statistic follows an approximate normal distribution
under the null hypothesis, that the population differences are centered at 0. The
mean µT and standard deviation T of T are given by:
Self-Instructional
258 Material
For a given level of significance a, the absolute sample Z should be greater Statistical Analysis
than the absolute Z/2 to reject the null hypothesis. For a one-sided upper tail test,
the null hypothesis is rejected if the sample Z is greater than Z and for a one-
sided lower tail test, the null hypothesis is rejected if sample Z is less than – Z. Let
us consider an example to illustrate the Wilcoxon-Rank test for a paired sample. NOTES
Example 12.11: A sample of 16 salesmen was selected in an organization and
their score on performance appraisal was noted. The salesmen were sent for a
three-week training programme and in the next appraisal, their scores were noted
again. The appraisal scores before and after the training are given below:
Use a 5 per cent level of significance to test the hypothesis that the training
has not caused any change in the performance appraisal score.
Solution:
H0 : There is no difference in the appraisal score because of training.
H1 : There is a difference in the appraisal score because of training.
The value of the T statistic can be worked out as follows:
Self-Instructional
Material 259
Statistical Analysis
NOTES
The test statistic Z is written as:
Self-Instructional
260 Material
which follows a 2 distribution with the k–1 degrees of freedom. Statistical Analysis
Use a 5 per cent level of significance to test the hypothesis that the amount
of wheat packaged by the three machines is the same.
Solution:
H0 : Amount of wheat packaged by the three machines is same.
H1 : Amount of wheat packaged by at least two machines is different.
Pool the elements of the different samples and rank them. These rankings
are shown below:
Self-Instructional
Material 261
Statistical Analysis Therefore,
NOTES
Non-parametric tests can often be applied to the nominal and ordinal data
that lack exact or comparable numerical values. NOTES
Non-parametric tests involve very simple computations compared to the
corresponding parametric tests.
Non-parametric methods are less powerful than parametric tests when the
basic assumptions of parametric tests are valid. Therefore, there is more
risk of accepting a false hypothesis and thus committing a type II error.
For the use of a chi-square test, the data is required in the form of
frequencies. The data expressed in percentages or proportion can also be
used, provided it could be converted into frequencies. The majority of the
applications of chi-square (2) are with the discrete data. The test could
also be applied to continuous data, provided it is reduced to certain
categories and tabulated in such a way that the chi-square may be applied.
There are many applications of a chi-square test. Some of them are explained
below:
o A chi-square test for the goodness of fit.
o A chi-square test for the independence of variables.
o A chi-square test for the equality of more than two population
proportions.
One of the assumptions that are usually made by researchers is that a random
sample is drawn from the population. Most of the tests of significance based
upon the Z, t or F distribution make use of this assumption.
A run is defined as a sequence of like elements that are preceded and
followed by different elements or no elements at all.
The test discussed in Unit 11 is based upon the assumption that the samples
are drawn from a population having roughly the shape of a normal
distribution. This assumption gets violated, especially while using the non-
metric data (ordinal or nominal). In such situations, the standard tests can
be replaced by a non-parametric test.
Two-sample sign test is a non-parametric version of it. It is based upon the
sign of a pair of observations.
When testing the equality of more than two population means, one-way
ANOVA technique was used in Unit 11. One of the assumptions used in
ANOVA is that all the involved populations from where the samples are
taken are normally distributed. If this assumption does not hold true, the
F-statistic used in ANOVA becomes invalid. The normality assumptions
may not hold true when we are dealing with ordinal data or when the size of
the sample is very small.
Self-Instructional
Material 263
Statistical Analysis The Kruskal-Wallis test comes to our rescue during such situations. This is,
in fact, a non-parametric counterpart to the one-way ANOVA.
NOTES
12.5 KEY WORDS
Goodness of fit: It is a statistical test of how well the observed data supports
the assumption about the distribution of a population.
Run: It is defined as a sequence of like elements that are preceded and
followed by different elements or no elements at all.
ANOVA: It is the test used for testing differences among the means of the
populations by examining the amount of variation within each of these samples.
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
264 Material
Research Reporting
13.0 INTRODUCTION
A research study is a tedious task and calls for an exhaustive investigation on the
part of the researcher. This quite often leads to accumulation of bulk data obtained
from the research study. Even if the concerned study results in brilliant hypotheses
or a generalized theory, it is the responsibility of the researcher to format this bulk
study into a pattern that is easy to understand. This is where report writing comes
into play.
In this unit, you will learn the concept of a research proposal. You will also
learn about written and oral reports. An effective written report is a creative activity
that requires a lot of imagination. It requires a lot of effort and patience in order to
write a report. It is impossible to think of the progress of an organization without
effective written reports, since most of the business activities require sending letters,
reports, etc. Effective oral report involves verbal communication of an idea to a
listener. An oral report saves time and builds a healthy atmosphere in an organization
by bringing the employees closer to each other.
13.1 OBJECTIVES
Self-Instructional
Material 265
Research Reporting
13.2 STEPS INVOLVED IN WRITING A RESEARCH
REPORT AND CHARACTERISTICS OF A
GOOD RESEARCH REPORT
NOTES
A report can be defined as a written document, which presents information in a
specialized and concise manner. For example, a list of employees prepared by the
HR (Human Resource) department for salary distribution can be termed as a
report. In other words, a report is information presented in a logical and concise
manner.
There is a difference between report writing and other compositions because
a report is written in a very short and conventional form. A report should cover all
mandatory matters but nothing extra should be written. For writing a report, at
first the relevant data is collected and then it is presented in a concise and objective
manner. Then after successfully establishing the structure of the report, the formatting
features that improve the look and readability of the report are added.
Reports can be divided into different categories. The two main types of
reports are as follows:
Informational report
Interpretive report
Informational Report
The report that consists of a collection of data or facts and is written in an orderly
way is called an informational report. The main purpose of this type of report is to
present the information in its original form without any conclusion and
recommendation. Informational reports are further divided into four parts, which
are:
Inspection report: The report which shows the outcome of a product or
equipment to assure its proper functioning or to describe its quality is called
an inspection report. This type of report is mainly used in manufacturing
organizations.
Inventory report: The report which is made to keep the stock of various
things like furniture, equipment, stationery, utensils and other accessories is
called an inventory report.
Assessment report: These reports are made to maintain the database of
the employees in an organization. Generally, these reports are useful for the
HR department.
Performance report: The report which is made to measure the performance
of the employees in an organization for purposes like appraisal or promotion
are called performance reports.
Self-Instructional
266 Material
Interpretive Report Research Reporting
Interpretive reports are those reports which contain a collection of data with its
interpretation or any recommendation explicitly specified by the writer. This type
of report also includes data analysis and conclusions made by the report writer. NOTES
Writing interpretive reports is different from writing an informational report because
it contains different elements. The possible elements that can be used in the
interpretive reports are as follows:
Cover
Frontpiece
Title page
Copyright notice
Forwarding letter
Preface
Acknowledgements
Table of contents
List of illustrations
Abstract and summary
Introduction
Discussion
Conclusions
Recommendations
Appendices
List of references
Bibliography
Glossary
Index
Characteristics of a Good Report
Reports are used for various purposes by various departments of an organization.
Industries, governments, businesses and scientific projects, all resort to report
writing to collect information and keep track of their performance and progress.
The most important aspect of a report is to convey the information in clear-cut
terms. It should provide facts in a direct, straightforward and accurate style. In this
light, the characteristics of a good report can be classified under four heads, which
are as follows:
Language and style of the report
Structure of the report
Self-Instructional
Material 267
Research Reporting Presentation of the report
References in the report
Each of these aspects of report writing needs to be given due attention, as
NOTES they are interrelated to each other. A report given with a lucid style but with very
less and hypothetical information is of no use to the reader. Similarly, the report
writer needs to avoid overcrowding of information that may make the reader feel
confused and lost in reading data, thereby losing its charm. A systematic scrutiny
of each of these aspects of a report is, therefore, necessary.
Language and style of report
A report must have a clear logical structure with clear indication of where the
ideas are leading. It should be able to make a good first impression. The presentation
of the report is very important. All reports must be written in good language, using
short sentences and correct grammar and spellings. The main points to be kept in
mind in this light are as follows:
Context and style
o Appropriate, informative title for the content of report
o Crisp, specific, unbiased writing with minimal jargon
o Adequate analysis of prior relevant research
Questions/hypotheses
o Clearly stated questions or hypotheses
o Thorough operational definitions of key concepts along with exact
wording or measurement of key variables
Research procedures
o Full and clear description of the research design
o Demographic profile of the participants/subjects
o Specific data gathering procedures
Data analysis
o Appropriate inferential statistics for sample or experimental data and
appropriate use of descriptive statistics
o Clear and reasonable interpretation of the statistical findings,
accompanied by effective tables and figures
Summary
o Fair assessment of the implications and limitations of the findings
o Effective commentary on the overall implications of the findings for theory
and/or policy
Self-Instructional
268 Material
Structure of report Research Reporting
Before you write a report, you should define the high level structure of the report.
Defining a clear logical structure will make the report easier to write and to read.
There are two types of report structures, which are listed as follows: NOTES
Report structure I: In general, the report writing structure comprises the
following sub-headings:
o Title Page
o Abstract
o Table of Contents
o Introduction
o Technical Detail and Results
o Discussion and Conclusions
o References
o Appendices
Report structure II: There is also a specific structure of report writing
pertaining to technical or scientific reports which is as follows:
o Introduction
o Background and Context
o Technical Details
o Results
o Discussion and Conclusion
Order of writing: The order of writing must be as follows:
o Start with the technical chapters/sections.
o Follow with the discussion.
o Finally, write the conclusions, introduction and abstract, if you are including
any.
Appendix: The appendix should contain the following:
o Material that suits or goes well with the flow of the main report but
cannot be included in the main text of the report either because it is too
long or is not essential reading. For example, lists of parameter values,
etc.
o Bibliography, i.e., list of all the sources of material, you referred to in
your report.
Presentation of report
As stated earlier, mere data overloading or just a lucid style of writing may not be
a plus point to good report writing. Both the aspects need to be given due
Self-Instructional
Material 269
Research Reporting consideration, so that they interact to give a simple, easy-to-read and
comprehensive type of report. Same goes with the presentation of the contents
of the report. Printing mistakes, informal use of font size and style can distract
the attention of the reader. On the other hand, effective use of tables and figures
NOTES for better understanding of data and writing its conclusions facilitate easy
comprehension. The main points of focus, where due attention is required on part
of the report writer are as follows:
Capitals: This requires taking care of the following aspects:
o Using capitals only for proper nouns, place names, organization names,
etc.
o Defining acronyms at the first point of usage. For example, Incorporated
(Inc).
o Using bold, italics or underlines for emphasis, instead of capitals.
Headings: The basic points to be kept in mind for headings are as follows:
o Differentiate headings from the rest of the text using different fonts, bold,
italics or underlines.
o Maintain consistency in formatting headings using predefined styles.
o Avoid headings beyond three levels.
Tables, figures and equations: In general, certain formatting standards
are pursued while giving tables and figures that are as follows:
o Descriptive labelling of all tables at the top with reference in the text.
o All figures must be labelled descriptively at the top and must be referenced
in the text.
o All equations must be numbered consecutively.
General presentation: The presentation of research follows the following
general points:
o Sheets should be plain like white A4 size, printed in one side only.
o Text should be justified on both sides and leave a blank line between
paragraphs.
o A staple in the top right hand corner is sufficient for most of the reports.
References in the report: Several report types like scientific, engineering,
technical and census reports contain either original writing or text adopted
from previous work. As such, a report writer should be careful and avoid
the violation of copyright laws and plagiarism. The necessary rule of thumb
in this regard can be stated as follows:
o Citations and referencing
– A citation is the acknowledgement in your writing of the work of other
authors and includes paraphrasing and making direct quotes.
Self-Instructional
270 Material
– Unless citation is very necessary, you should write the material in your Research Reporting
own words. This shows that you understand what you have read and
know how to apply it, to your own context.
– Direct quotes should be used sparingly.
NOTES
o Direct quotes
– Short direct quotes: These need to be placed between quotation
marks. For example, Rosenfield defines a cluster as a ‘geographically
bounded concentration of similar, related or complementary businesses,
with active channels for business transactions, communications and
dialogue that share specialized infrastructure, common opportunities
and threats’. This shows clearly that the words being used are not
your own words.
– Longer direct quotes: There are occasions when it is useful to include
longer direct quotes. If you are quoting more than about 40 words,
you should again use quotation marks but also indent the text. For
example, the sustainability of higher value added industry is grounded
in the diminishing significance of cost structures. At the level of the
European Union, a weak capacity to innovate has been identified as
an innovation, in the sense of product, process, and organizational
innovation, accounts for a very large amount, perhaps 80–90 per cent
of the growth in productivity in advanced economies.
Mechanics of Writing a Report
There are several mechanics of writing a report, which are strictly followed for
preparing technical reports. The following points should be considered for writing
a technical report:
Size and physical design: The manuscript, if handwritten, should be in
black or blue ink and on unruled paper of 8½" × 11" size. A margin of at
least one-and-half inches is set at the left side and half inch at the right side
of the paper. The top and bottom margins should be of one inch each. If the
manuscript is to be typed, then all typing should be double spaced and on
one side of the paper, except for the insertion of long quotations.
Layout: According to the objective and nature of the research, the layout
of the report should be decided and followed in a proper manner.
Quotations: Quotations should be punctuated with quotation marks and
double spaces, forming an immediate part of the text. However, if a quotation
is too lengthy, then it should be single spaced and indented at least half an
inch to the right of the normal text margin.
Footnotes: Footnotes are meant for cross-references. They are placed at
the bottom of the page, separated from the textual material by a space of
half an inch as a line that is around one-and-a-half inches long. Footnotes
Self-Instructional
Material 271
Research Reporting are always typed in a single space, though they are divided from one another
by double space.
Documentation style: The first footnote reference to any given work
should be complete, giving all essential facts about the edition used. Such
NOTES
footnotes follow a general sequence and order:
o In case of the single volume reference:
– Author’s name in normal order
– Title of work, underlined to indicate italics
– Place and date of publication
– Page number reference
For example:
John Gassner, Masters of the Drama. New York: Dover Publications,
Inc.1954, p.315.
o In case of a multi-volume reference:
– Author’s name in the normal order
– Title of work, underlined to indicate italics
– Place and date of publication
– Number of the volume
– Page number reference
For example:
George Birkbeck Hill, Life Of Johnson. Whitefish, June 2004, Volume
2, p.124.
o In case of works arranged alphabetically:
– For works arranged alphabetically such as encyclopaedias and
dictionaries, page reference is usually not needed. In such cases, order
is illustrated according to the names of the topics.
– Name of the encyclopaedia
– Number of editions
For example:
‘Salamanca’, Encyclopaedia Britannica, 14th Edition.
o In case of periodicals reference:
– Name of the author in normal order
– Title of article, in quotation marks
– Name of the periodical, underlined to indicate italics
– Volume number
– Date of issuance
– Pagination
Self-Instructional
272 Material
For example: Research Reporting
part of the research process. No matter how well designed the research study is,
it is of little value, unless communicated effectively to others in the form of a research
report. Moreover, if the report is confusing or poorly written, then the time and
NOTES
effort spent on gathering and analysing data would be wasted. It is therefore,
essential to summarize and communicate the result to the management of an
organization with the help of an understandable and logical research report.
Research reports are helpful during the research study, in the sense that they
facilitate maintenance of vast data in a logical way. Thus, in case the researcher
experiences any difficulty during the course of the study, it becomes easier to refer
to the contents of the report to get the relevant data. Research report writing
essentially involves systematic arrangement of data. This helps in discovering flaws
in reasoning, which may have been missed earlier while conducting a research.
Format of Research Report
The layout of the research report is of utmost importance because the reader
should be able to grasp logically, what has been said and not feel lost in the bulk
findings mentioned in the research. This requires preparing a proper layout of the
report. Report layout means allotting the research findings in a comprehensible
format. The layout should contain the following points:
Preliminary pages: In the preliminary pages, the report should carry a
‘title’ and a ‘date’, followed by acknowledgements in the form of ‘Preface’
or ‘Foreword’. The ‘Table of Contents’ should come next, followed by a
‘list of tables and illustrations’. This entails the reader to an easy reading
and quick location of the required information.
Main text: The main text comprises the complete outline of the research
report with all the details. The title of the research study is repeated at the
top of the first page of the main text and then followed with the other details
on the pages numbered consecutively, beginning with the second page. The
main text can be classified into the following sections:
o Introduction: The purpose of introduction is to introduce the research
projects to the readers. It should clearly state the objectives of research,
i.e., it should make clear, why the problem was considered worth
investigating. A brief summary of other relevant research can be included
as well, to enable the reader to see the present study in that context.
o Methodology used for performing the study: The introduction should
contain answers to questions like how was the study carried out, what
was the basic design, what were the experimental directions, what
questions were asked in the questionnaires used, etc. Besides this, the
scope and limitations of the study must be marked out.
Self-Instructional
Material 275
Research Reporting o Statement of findings and recommendations: The research report
should comprise a statement of findings and recommendations in a non-
technical language so that it is easily comprehensible.
report.
It should be attractive, neat and clean, whether handwritten or typed.
The report writer should be careful about the possessive form of the word NOTES
‘it is’ with ‘it’s’. The correct possessive form of ‘it’s’ is ‘its’. The use of ‘it is’
is the contractive form of ‘it is’.
A report should not have contractions. Examples are ‘didn’t’ or ‘it’s’. In
report writing, it is best to use the non-contractive form. Hence, the examples
would be replaced by ‘did not’ and ‘it is’. Using ‘Figure’ instead of ‘Fig.’
and ‘Table’ instead of ‘Tab.’ will spare the reader of having to translate the
abbreviations, while reading. If abbreviations are used, use them consistently
throughout the report. For example, do not switch between ‘versus’ and
‘vs’.
It is advisable to avoid using the word ‘very’ and other such words that try
to embellish a description. They do not add any extra meaning and, therefore,
should be dropped.
Repetition hampers lucidity. The report writer must avoid repeating the same
word more than once within a sentence.
When using the words ‘this’ or ‘these’, it must be clear to the reader as to
what is being referred to. This reduces ambiguity in the writing and helps to
tie sentences together.
Do not use the word ‘they’ to refer to a singular person. You can either
rewrite the sentence to avoid needing such a reference or use the singular
‘he or she.’
Self-Instructional
Material 277
Research Reporting 2. The order of writing must be as follows:
Start with the technical chapters/sections
Follow with the discussion
NOTES Finally, write the conclusions, introduction and abstract, if you are including
any.
3. The ‘list of tables and illustrations’ are the last section of the preliminary
pages.
4. The introduction of a research should clearly state the objectives of research,
i.e., it should make clear, why the problem was considered worth
investigating. A brief summary of other relevant research can be included as
well, to enable the reader to see the present study in that context.
5. The end of the research report should consist of appendices, listed in respect
of all technical data such as questionnaires, sample information and
mathematical derivations. The bibliography of the referred sources and an
index should also be given.
13.4 SUMMARY
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Self-Instructional
Material 279
Research Reporting Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
NOTES
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
280 Material
Applications of Computer
COMPUTER IN
NOTES
EDUCATIONAL RESEARCH
Structure
14.0 Introduction
14.1 Objectives
14.2 Uses of Computer in Data Analysis
14.3 Application of MS-Office
14.4 Basics of MS-Word
14.5 MS-Excel
14.6 MS-PowerPoint
14.7 Application of these Softwares for Documentation and Making Reports
14.8 Use of SPSS and Other Statistical Softwares
14.9 Answers to Check Your Progress Questions
14.10 Summary
14.11 Key Words
14.12 Self Assessment Questions and Exercises
14.13 Further Readings
14.0 INTRODUCTION
14.1 OBJECTIVES
Computers are useful in various ways. With the increasing availability of more
NOTES complex and dynamic operating systems, the primary use of a computer is only
limited to the imagination and technical know-how of the user. Everything from
your cell phone, DVD player, to your TV has some sort of microprocessor in it,
giving it computer-like abilities. Computers have the ability to help the society by
various means. They have a number of applications in science as well. In space
aeronautics, computers are used in space shuttles for data collection as well as
control of flights. In the medical industry, computers are being used in conjunction
with robotics to create a new breed of machines that can perform operations on a
minimally invasive scale, thereby, increasing the patient’s survival rate and reducing
the healing time. Computers are also used in agriculture for controlling complex
irrigation systems and sensors that detect soil pH among others things, to give the
crops a higher yield and faster grow times.
The various applications of computers are as follows:
Word processing: Word Processing software automatically corrects spelling
and grammar mistakes. If you want some content of your document to be
repeated, you do not have to type it each time. You can use the copy and
paste features. Images can also be added to your document.
Internet: It is a network of almost all the computers in the world. You can
browse through much more information than you could do in a library. That
is because computers can store enormous amounts of information. You can
also have very fast and convenient access to information. Through e-mail,
you can communicate with a person sitting thousands of miles away in
seconds. The chat software enables you to chat with another person on a
real-time basis. Videoconferencing tools are becoming readily available to
the common man.
Digital video or audio composition: Audio or video composition and
editing have been made much easier by computers. It no longer costs
thousands of dollars of equipment to compose music or make a film. Graphics
engineers can use computers to generate short or full-length films or even
create three-dimensional models. Anybody owning a computer can now
enter the field of media production. Special effects in science fiction and
action movies are created using computers.
Desktop publishing: Page layouts for books can also be created on your
personal computer.
Medicine and health care: Software is used in magnetic resonance imaging
to examine the internal organs of the human body and also used for
performing surgery. Computers are used to store patient data.
Mathematical calculations: Computers have computing speeds of over
Self-Instructional
282 Material
a million calculations per second, using which we can perform various Applications of Computer
in Educational Research
mathematical calculations.
Banks: All financial transactions these days are done by computer software.
They provide security, speed and convenience.
NOTES
Travel: One can book air tickets or railway tickets and make hotel
reservations online.
Communication: Software is widely used in this field through which you
can interact with people around the world.
Military: There is software embedded in almost every weapon. Military
software is used for controlling flights and for marking target in ballistic
missiles. Software is used to control access to atomic bombs.
E-learning: For a student, it is easier to learn from e-learning software
instead of a book.
Examinations: You can give online exams and get instant results.
Certificates: Different types of certificates can be generated and it is very
easy to create and change layouts.
ATM machines: The computer software authenticates the user and
dispenses cash for banks.
Marriage: There are matrimonial sites through which one can search for a
suitable groom or bride.
News: There are many websites through which you can read the latest or
old news.
Planning and management: Software can be used to store contact
information, generate plans and schedule appointments and deadlines.
Plagiarism: Software can examine content for plagiarism.
Sports: It is used for making umpiring decisions. There are simulation
software using which a sportsperson can practice his skills. Computers are
also to identify flaws in technique.
Airplanes: Pilots train on software, which simulates flying.
Weather analysis: Supercomputers are used to analyse and predict
weather.
Research: Computers are widely used for research purposes in various
fields such as follows:
o Network-attached storage (Linux distribution named FreeNAS)
o Media Server (Hewlett-Packard makes a dedicated version)
o Graphics design (Adobe is the forefront in design software)
o Architectural design (AutoCAD/CAM)
o Online banking (savings, loans, insurance, credit, mutual funds, etc.)
Self-Instructional
Material 283
Applications of Computer o Gaming (computer 3D games, etc.)
in Educational Research
o Social networking (Myspace, Facebook, Twitter)
o Knowledge sharing (WikiAnswers, Wikipedia, Lifehacker, Gizmodo,
NOTES etc.).
Computers are not being used extensively for data analysis. It helps in
the following actions:
Locating segments of data in the form of words or phrases
Storing, annotating and retrieving texts
Labeling or naming data
Organizing and Sorting the data
Drawing figures, maps, tables or charts
When it comes to quantitative data, computers now have different software
which helps in the input of data and its calculation the computers which eliminates
the efforts of calculation and such tasks. In the succeeding sections, you will learn
about the application of different software for documentation and report making.
Self-Instructional
284 Material
as well as sharing a file. This is also used for closing applications. A color schemes Applications of Computer
in Educational Research
can also be chosen by users for the interface.
Ribbon: Ribbon is a panel that houses a fixed arrangement of few icons and
command buttons. This creates organizations for commands that form a set of
NOTES
tabs that group every relevant command. This can not be customized. Every
application contains a different set of tabs that tells about its functions.
Contextual Tabs: There are certain tabs that show appearance only after making
selection of certain objects. These tabs are known as Contextual Tabs. Such
tabs focus on functions specific to the selected objects only. If you select a picture
then Pictures tab is shown and this gives all options to work with pictures.
Mini Toolbar: This is a new inclusion in Office 2007. It appears automatically as
a context menu when you select a text. This design provides easy access to
formatting commands that are used repeatedly without making use of right button
of the mouse as is done in older versions.
Super Tool Tips: Super tool tips are also known as screen tips. This is capable
of housing formatted text as well as images. These show detailed descriptions
about most buttons and their functions.
Quick Access Toolbar: This tool bar is inside the title bar and is a repository of
functions that are most commonly used. This toolbar can be customized. Any
command can be included to this toolbar and this also includes commands that are
not available in Ribbon and macro.
Microsoft Office Word 2007 enables you to create formal documents by providing
a broad set of tools for crafting and formatting your documents in a new interface.
Rich productive review, remarking content with comments and comparison
capabilities help you to manage feedback from colleagues. Significant sources of
business information stay connected by advanced and improved data integration.
Working in Compatibility Mode
When you open Word 97-2003 document in Microsoft Office Word 2007,
Compatibility Mode is turned on and you see Compatibility Mode in the title bar
of the document window. In Compatibility Mode, you can open, edit and save
Word 97-2003 documents but you will not be able to use any of the new Office
MS Word 2007 features.
Self-Instructional
Material 285
Applications of Computer Compatibility Checker
in Educational Research
The Compatibility Checker lists elements in your document that are not supported
or will work differently in Word 97-2003 format. In the Compatibility Checker,
NOTES you can review a summary of elements that behave differently in previous versions
of Word and then either click Continue to save the document in Word 97-2003
format or click Cancel. Some of these features will be permanently changed and
will not be converted to Microsoft MS Office Word 2007 elements if you convert
the document to Office Word 2007 format.
Citations and Bibliographies: Citations and bibliographies will be
converted to static text.
Content Controls: Content controls will be converted to static content.
Embedded Objects: An embedded object in this document was created
in Microsoft Office Excel 2007. You will not be able to edit the object in
earlier versions of Word.
Equations: Equations will be converted to images. You will not be able to
edit the equations until the document is converted to a new file format.
SmartArt Graphics: SmartArt graphics will be converted into a single
object that cannot be edited in previous versions of Word.
Tabs: Alignment tabs will be converted to traditional tabs.
Text Boxes: Some text box positioning will change.
Tracked Moves: Tracked moves will be converted.
Creating and Editing Document in Microsoft Word 2007
The Start button in the lower-left corner of your screen gives you access to all the
programs on your PC and also to MS Word 2007. To start Microsoft Word
2007, click on the Start button and select All Programs. To open this window,
you will need to perform the following steps:
Self-Instructional
286 Material
Click on the Start button and select Microsoft Office from Programs. Applications of Computer
in Educational Research
Select Microsoft Office Word 2007.
The user interface of Microsoft Word 2007 is shown in the screen below.
NOTES
Menus
When you explore Microsoft Word 2007, you will notice the new look of the
menu bar. Three new features help you to work with MS Word 2007, namely the
Microsoft Office Button, the Quick Access Toolbar and the Ribbon which contain
various functions.
The Microsoft Office Button
The Microsoft Office button is located in the upper-left corner of the MS Word
2007 window. A menu appears when you click on this button. This menu helps in
creating a new document or file, opening an existing document or file, saving a
document or file, printing a document or file, sending the document or file via fax
or e-mail, etc.
Self-Instructional
Material 287
Applications of Computer
in Educational Research
You can customize this toolbar as per your requirements by clicking on the
NOTES
expansion button as shown below.
More items can be added to the quick access toolbar by right clicking on
the item which you want to add in the Office Button or the Ribbon and then
clicking on Add to Quick Access Toolbar as shown below:
The Ribbon
The Ribbon is positioned at the top of the screen of the Word window. It includes
seven tabs, namely Home, Insert, Page Layout, References, Mailings,
Review, View and Add-Ins as shown in screen below. Each tab contains various
new and advanced features of Word.
Self-Instructional
288 Material
Mailings: Create, Start Mail Merge, Write & Insert Fields, Preview Results Applications of Computer
in Educational Research
and Finish
Review: Proofing, Comments, Tracking, Changes, Compare and Protect
View: Document Views, Show/Hide, Zoom, Window and Macros NOTES
Add-Ins: PDF Transformer or any new Add-In program
The Title Bar
The Title bar is next to the Quick Access Toolbar. It displays the title of the current
document which is in use. The first new document in Word is named as ‘Document1’
as shown below. When you open more new documents, Word automatically names
them as Document2, Document3, Document4, etc. sequentially. The document
can be saved by giving it a proper file name as per the user’s choice.
Word provides various ways through which a user can create a new document,
open an existing document and save a document in Word. To create a new document,
click on the Microsoft Office Button and then click on New or press CTRL+N
on the keyboard.
You will see that when you click on the Microsoft Office Button and then
click on New, Word provides number of choices about the types of documents
you can create. Select and click on Blank from the list if you want to create a
blank document.
You can create a document using the option Installed Templates. Select
any one of the templates as per your requirement. You can also browse other
options through the list of choices that appear on the left.
Self-Instructional
Material 289
Applications of Computer
in Educational Research
NOTES
Saving a Document
To save a document, the following are the options available to opt from according
to the user’s choice:
Click on the Microsoft Office Button and click on the Save option.
The alternate option is to press CTRL+S.
Another option is to click on the File icon in the Quick Access Toolbar and
click on the Save.
NOTES
Renaming a Document
To rename an existing document, you need to perform the following steps:
Click on the Office Button and locate the document you want to
rename.
Click on the Save As option and then right click on the document name
with the mouse and select Rename from the shortcut menu.
Type the new name for the document and press the ENTER key.
Working on Multiple Documents
Multiple documents can be simultaneously opened when you need to type or edit
multiple documents at once. All the documents as opened will be listed in the View tab
of the Ribbon when you will click on Switch Windows. The current document has a
checkmark beside the file name. You can select a different document as opened by
simply clicking on the tab.
Self-Instructional
Material 291
Applications of Computer
in Educational Research
NOTES
Self-Instructional
292 Material
Closing a Document Applications of Computer
in Educational Research
To close a document, the following steps need to be performed:
NOTES
In MS Word 2007, the commands you need to carry out any move or copy
operation are located on the Home tab within the Clipboard. Following are the
keyboard shortcuts that are also helpful when moving through the text of a
Self-Instructional document:
294 Material
Applications of Computer
Move Action Keystroke in Educational Research
Beginning of the line HOME
End of the line END
Top of the document CTRL+HOME NOTES
End of the document CTRL+END
Copy the selected text. Select Paste option and place the mouse pointer
where you want to paste the selected material. Click on Paste option. The copied
Self-Instructional
Material 295
Applications of Computer and cut text will be stored in Clipboard application as shown in the following
in Educational Research
screen.
NOTES
For cut operation, you need to select the text. Click on Cut option in shortcut
key. Place the mouse pointer where you want to paste the cut text. Click on Paste
option in shortcut key.
Finding and Replacing Text
The Find and Replace option can be accessed either by selecting CTRL+F or
CTRL+H menu by pressing key combinations for Find and Replace. After
choosing the Find or Replace option, you will get the following screen:
Self-Instructional
296 Material
The match case provides you to find and replace the word as uppercase or Applications of Computer
in Educational Research
lowercase. For example, if you check on Match case box and type the word in
capital as ‘TOP’, a dialog box appears with a message ‘The search item was
not found’.
NOTES
If you remove the check box, the found word is replaced by defined word
as follows:
You can also search the items by using wild card (*) as shown in screen
below:
Insert/Delete Text
In Microsoft Word 2007, you can create documents by typing them, for example,
if you want to create a report, you open Microsoft Word 2007 and then start to
type. You do not have to do anything when your text reaches the end of a line and
you want to move to a new line, Microsoft Word 2007 automatically moves your
text to a new line. If you want to start a new paragraph press ENTER key.
Self-Instructional
Material 297
Applications of Computer Microsoft Word 2007 creates a blank line to indicate the start of a new paragraph.
in Educational Research
To capitalize, hold down the SHIFT key while typing the letter you want to
capitalize. If you make a mistake, you can delete what you typed and then type
your correction. You can use the Backspace key to delete. Each time you press
NOTES the Backspace key, Microsoft Word 2007 deletes the character that precedes
the insertion point. The insertion point is the point at which your mouse pointer is
located. You can also delete text by using the Delete [DEL] key. First, you select
the text you want to delete and then you press the Delete key.
Creating Template
A Microsoft Office Word 2007 template contains sample content, formatting or
objects that can be used to quickly and easily create a new document. To create
a template, create a normal document and save it by selecting from the Office
button Save As Word Template as shown below. The file extension
assigned will be either .dotx if the template has no macros or code or .dotm if the
template contains macros or code.
The templates contain the formatting and layout for the documents to be
created quickly and easily. Whenever you create a new document using a template
the predefined settings will be automatically added to it. To create and save a
template, follow the given steps:
Open the document that you want to save as a template or create a new
document.
Click on the Microsoft Office button and then click on Save As.
In ‘Save a copy of the document’ dialog box select Word Template.
Browse to the location where you want to save the template and then type
a file name for the template.
Click on Save.
To open a document based on the new template, browse to the location where
the template is saved and then double-click the file name. You can save a document
as a template at any time and also update or modify the template as per your
requirements. You can protect the template from other users if you do not want it
Self-Instructional
298 Material
to be changed by someone else. Set the file properties to “read-only” instead of Applications of Computer
in Educational Research
“read-write” by following the given steps:
Browse to the location where the template is saved.
Right-click the template file name and then click on Properties. NOTES
The Property window will appear. On the General tab next to Attributes,
select the Read-only check box and then click on OK.
Creating Tables
Tables are used to display the data in a tabular format.
Create a Table
To create a table, follow the given steps:
Place the cursor on the page where you want the new table
Click on the Insert Tab of the Ribbon
Click on the Tables Button on the Tables Group. You can create a table
using any one of the following four ways:
Highlight the number of row and columns
Click on Insert Table and enter the number of rows and columns
Click on the Draw Table, create your table by clicking and entering the
rows and columns
Click on Quick Tables and select a table
Inserting Rows/Columns
Add a row above or below: Follow the given steps to add a row:
Click in a cell above or below where you want to add a row.
Under Table Tools on the Layout tab to add a row above the cell click on
Insert Above in the Rows and Columns group.
Under Table Tools on the Layout tab to add a row below the cell click on
Insert Below in the Rows and Columns group.
Self-Instructional
Material 299
Applications of Computer Add a column to the left or right: Follow the given steps to add a column:
in Educational Research
Click in a cell to the left or right of where you want to add a column.
Under Table Tools on the Layout tab to add a column to the left of the
NOTES cell click on Insert Left in the Rows and Columns group.
Under Table Tools on the Layout tab to add a column to the right of the
cell click on Insert Right in the Rows and Columns group.
Merging Rows/Columns
You can combine two or more cells in the same row or column into a single cell.
For example, you can merge several cells horizontally to create a table heading
that spans several columns.
Select the cells you want to merge.
On the Table menu click on Merge Cells.
Moving a Table
Word 2007 allows you to use the mouse to move entire table within your document.
You can do this by using techniques similar to those you use to move graphics
around in a document. Position the mouse over your table. At the upper-left corner
of the table appears a small icon. This icon looks like a square with a four-headed
arrow inside it. Click and drag this icon where you want to move the table. When
you release the mouse button the table is repositioned where you dragged the
mouse.
Entering Data in a Table: To enter data in a table, place the cursor in the
cell where you want to enter the data or information. Start typing to enter the data.
Inserting Symbols and Special Characters
Word 2007 permits you to insert special characters, symbols, pictures, illustrations,
etc. Special characters are punctuation, spacing or typographical characters that
are not generally available on the standard keyboard. To insert symbols and special
characters, follow the given steps:
Place your cursor in the document where you want the symbol
Click on the Insert Tab on the Ribbon
Click on the Symbol button on the Symbols Group
Select the proper symbol.
Self-Instructional
300 Material
If you want more symbols then click on More Symbols to display the following Applications of Computer
in Educational Research
dialog box displaying list of various symbols in various fonts.
NOTES
Equations
Word 2007 also permits you to insert mathematical equations. To use the
mathematical equations tool, follow the given steps:
Place your cursor in the document where you want the equation
Click on the Insert Tab on the Ribbon
14.5 MS-EXCEL
MS Excel 2007 files are referred as spreadsheets. This is a generic term, which
sometimes means a workbook (file) and sometimes means a worksheet (a page
within the file). Data files created with MS Excel 2007 are called workbooks. MS
Excel 2007 files by default contain three blank worksheets. This gives you the
Self-Instructional
Material 301
Applications of Computer flexibility to store related data in different locations within the same file. More
in Educational Research
worksheets can be added and unwanted worksheets can be deleted as per the
user requirement. Thus, MS Excel 2007 is a powerful and most extensively used
tool as spreadsheet application which allows you to store, organize and analyse
NOTES numerical, graphic and text data. Spreadsheets allow information to be organized
in rows and tables, and can be analysed using various mathematical, trigonometric,
text, logical, date and time functions. The number of rows is now 1,048,576 (220)
and columns is 16,384 (214). Microsoft Excel 2007 has the basic features of all
spreadsheets, using a grid of cells arranged in numbered rows and letter named
columns to organize data manipulations. Further with Microsoft Excel 2007 you
can analyse, manage and share information quickly and easily to formulate more
knowledgeable decisions. With the new user interface, rich data visualization and
PivotTable views, professional looking charts can be created easily.
The advanced features in Microsoft Excel 2007 include Office themes, more
styles, rich conditional formatting, easy formula writing, Sort & Filter, data validation,
worksheet and workbook protection, Goal Seek, Scenario, PivotTable and
PivotChart. Goal Seek and Scenario are part of What-If Analysis tools. MS Excel
2007 supports charts, graphs or histograms generated from specified groups of
cells. The generated graphic component can either be embedded within the current
sheet or added as a separate object. OLE or Object Linking and Embedding
allow a Windows application to format or calculate data. This may acquire the
form of ‘embedding’ where an application uses another to handle the task, for
example a MS PowerPoint presentation can be embedded in an MS Excel 2007
spreadsheet or vice versa.
You can create a spreadsheet using various formatting and editing features
given in the Ribbon panel. Thus, you can perform calculations using the functions
given in MS Excel 2007 and can sort and filter data as per your requirement. You
can create graphs in the worksheet, insert illustrations, Clip Art, SmartArt, Shapes
and pictures to enhance your worksheet. You can freeze and unfreeze rows and
columns and can also hide and unhide any worksheet.
Worksheet, Workbook and Workspace
A Microsoft Excel 2007 file in which you can enter and store related data is
known as a workbook. A workbook is also identified as a spreadsheet that is a
group of cells on a single sheet where you in fact keep and operate data. Every
worksheet consists of columns and rows. The columns are lettered A to Z and
then continue with AA, AB, AC, and so on. The rows are numbered from1 to
1,048,576. The number of rows and columns that you can hold in your worksheet
is restricted by computer memory and your system resource.
Self-Instructional
302 Material
Applications of Computer
in Educational Research
NOTES
Cell address is the combination of a column letter and a row number. For
example, if a cell is located in the upper left corner of the worksheet, which is A1,
this means that it is located in column A and row 1. Similarly, cell E10 is situated in
column E on row 10. The data can be entered into the cells present on the worksheet.
N-number of worksheets can be present in a workbook. To work with workspace
you have to perform the following steps:
Click the Office button. Click Open.
Self-Instructional
Material 303
Applications of Computer
in Educational Research
NOTES
Getting Started
As soon as you start MS Excel 2007, you will see that its features are similar to
the previous versions. You will also see that there are many additional features
which help you to work with special effects. Three new features that are included
in MS Excel 2007 are the Microsoft Office Button, the Quick Access Toolbar
and the Ribbon.
Microsoft Office Button
The Microsoft Office Button performs various functions that were found in the
File menu of older versions of MS Excel 2007. This button permits you to create
a new workbook, open an existing workbook, save a workbook using Save and
Save As, Print, Send or Close a workbook.
Ribbon
The Ribbon is the panel at the top portion of the spreadsheet. It includes seven
tabs namely, Home, Insert, Page Layout, Formulas, Data, Review and View.
Add-ins is another option which is automatically displayed on the Ribbon when
you add any new application to the program. Each tab is a collection of features
designed to perform specific functions that you require while creating or editing
MS Excel 2007 spreadsheets.
Self-Instructional
304 Material
The frequently used features are displayed on the Ribbon. To view additional Applications of Computer
in Educational Research
features of each group, click on the arrow at the bottom right corner of each
group.
NOTES
You can also add more items to the Quick Access toolbar. To do this, right
click on any item in the Office Button or the Ribbon and then click on Add to
Quick Access Toolbar. A shortcut will be added there.
Self-Instructional
Material 305
Applications of Computer Mini Toolbar
in Educational Research
Mini Toolbar is a new feature in Microsoft Office 2007. This is a floating toolbar
and is displayed when you select text or right click any text. It displays the common
NOTES formatting tools, such as Bold, Italic, Fonts, Font Size and Font Color.
Click on MS Excel 2007 Options which you will get from Quick Access
Toolbar.
Popular
The Popular features helps you to personalize your work environment using the
Mini Toolbar, Color schemes, default options when creating new workbooks and
creating lists for sort and fill sequences. It also helps you to access the Live Preview
feature to preview how a feature affects the document as you hover over different
choices. The choices provide new font size, table style or cell style which can be
applied on a workbook as per requirement.
Self-Instructional
306 Material
Applications of Computer
in Educational Research
NOTES
Formulas
The Formulas feature permits you to modify the calculation options, to work with
formulas, error checking and error checking rules. Working with formulas provide
four check boxes which are R1C1 reference style, Formula AutoComplete, Use
table names in formulas and Use GetPivotData functions for PivotTable references
as shown in the given screen.
Proofing
The Proofing feature permits you to personalize the options for correcting words
and formats of your text. You will get AutoCorrect option in Proofing feature. You
can customize auto correction settings so that it will ignore certain words or errors
in a document via the Custom Dictionaries...
Self-Instructional
Material 307
Applications of Computer
in Educational Research
NOTES
Save
The Save feature permits you to personalize your workbook when saved. You
can also specify how often you want auto save to run and where to save the
workbooks.
Advanced
The Advanced feature permits you to specify the options for editing, copying,
pasting, printing as well as displaying formulas, calculations and other general
settings.
Self-Instructional
308 Material
Applications of Computer
in Educational Research
NOTES
Customize
Customize permits you to add specific features to the Quick Access Toolbar. It
adds the tools which you frequently use.
Self-Instructional
Material 309
Applications of Computer Click the Advanced category and scroll down up to the General section.
in Educational Research
In the box for ‘At startup, open all files in’, you might see the name of a
folder and its path. Clear the folder information from that box or go to that
folder and remove the unwanted files. Click OK to close the MS Excel
NOTES
2007 Options dialog box.
Entering Information in a Worksheet
To enter information in a worksheet, you need to open an empty workbook and
enter the data as shown in the screen, below:
Self-Instructional
310 Material
calculation. Select the worksheets that you want to move or copy as shown in the Applications of Computer
in Educational Research
screen below.
NOTES
To move to the next or previous sheet tab, you can also press CTRL + Pg
Up or CTRL + Pg Dn. On the Home tab, in the Cells group, click Format and
then under Organize Sheets, click Move or Copy Sheet.
You can also right click a selected sheet tab and then click Move or Copy. In the
Move or Copy dialog box, in the Before sheet list, do one of the following:
Click the sheet before which you want to insert the moved or copied sheets.
Click move to end to insert the moved or copied sheets after the last
sheet in the workbook and before the Insert Worksheet tab.
To copy the sheets instead of moving them, in the Move or Copy dialog
box, select the Create a copy check box.
Saving a Workbook
To save a workbook, you have two options, Save and Save As. To save a
document, follow the given steps:
Click on the Microsoft Office Button.
Click on Save.
You can also use the Save As feature to save the workbook with a different
name or to save it as earlier versions of MS Excel 2007. The older versions of
MS Excel 2007 cannot be opened in an MS Excel 2007 worksheet unless you
Self-Instructional
Material 311
Applications of Computer save it as an MS Excel 97-2003 Format. To use the Save As feature, follow the
in Educational Research
given steps:
Click on the Microsoft Office button.
Click on Save As.
NOTES
Give a name for the workbook.
In the Save as Type box, select Excel 97-2003 workbook.
Self-Instructional
312 Material
Quitting From MS Excel 2007 Applications of Computer
in Educational Research
To quit from MS Excel 2007, click on Microsoft Office Button and then select
Exit MS Excel 2007 button.
You will quit from MS Excel 2007. NOTES
14.6 MS-POWERPOINT
PowerPoint helps in using charts, diagrams, pictures and animations for the purpose
of creating effective presentation slides. The main feature that separates
PowerPoint 2007 from PowerPoint 2003 is that in PowerPoint 2007 file is saved
with a .ppt and .pptx extension. When the PowerPoint slides are saved as .pptx,
Windows 2003 is unable to open the file.
Self-Instructional
Material 313
Applications of Computer Microsoft Office Button
in Educational Research
The Microsoft Office Button performs all the functions of the ‘File’ menu of the
older versions of PowerPoint. It helps you to create a new presentation, open an
NOTES existing presentation, save a presentation, save a presentation with a new name
using the ‘Save As’ option, print a presentation, send a presentation and close a
presentation.
Ribbon
Ribbon refers to the strip of buttons that resides on top of the main Window. The
standard Ribbon includes the Home tab, the Insert tab, the Design tab, the
Animation tab, the Slide Show tab, the Review tab and the View tab.
Design
The Design option is accessed on the Ribbon. This option facilitates the choice of
colors, background styles, fonts, page setup, slide orientation, etc.
Home
The Home option is the most commonly used option which by users. It helps in
creating new slides. The ‘slides’ option provides you to insert new slides. You can
adjust the layout of slides, reset and set default slides. The paragraphs can be
aligned and specified in form of bulleted ornumbered lists. The drawing and editing
tools help in editing the text and figures.
Insert
The insert option is available for the purpose of adding tables, illustrations, links,
text and media clips. WordArt, header, footer, text, movie and sound can also be
inserted in the slides. The tables can be inserted or imported from MS Excel.
Illustrations can be in form of Clip Art files, photo albums, pictures, Smart Art,
shapes and charts. You can insert a link using the hyperlink tool to navigate the
Self-Instructional
314 Material
corresponding presentations. The ‘insert text box’ option provides the orientation Applications of Computer
in Educational Research
and location of the words along with the insertion of date, time, symbols, slide
numbers and embedded objects.
NOTES
Animation
The animation option contains preview, custom animation and various transition
settings that can be applied to the specified or all slides in the presentation. The
slide show transition can be set at mouse clicks or ‘automatically-after -seconds’
options. You can preview the slide show to view the proper effect and the mode of
presentations. Various objects, such as images, text and embedded objects can
be added on the slides. The various transitions available for slide shows are wipes,
fades and dissolve, random, strips and bars, push and cover, etc.
Slide Show
The slide show option helps in setting up the start slide show (either from the
starting or from-and-to specific slide numbers) and to record narration. It also has
the option to monitor the screen resolution by providing separate views of the
slide show.
Review
The content of the slides can be reviewed and modified by using the spell check,
research, adding comments, etc. This makes the presentation flawless. Proofing
provides the facility of text proofing by scanning the online research references,
finding synonyms and converting the text to other languages in totality. This option
provides the comment facility that enables the addition or modification of a comment
for a particular slide or the content of a slide. The protect option restricts usage by
unauthorized users. This option is helpful for slide show share with a network
drive if you collaborate with other users.
Self-Instructional
Material 315
Applications of Computer
in Educational Research
NOTES
View
The view option enables the presentation to be viewed in different ways, such as
normal view, notes view, handout view, printout view and screen view, show/hide
grid lines, rulers and tools, zoom in and zoom out facility and also includes the
color/grayscale view whether the slides should appear in color or black and white.
The window tab arranges the windows of the current working slides and macros
includes complex tasks that get activated after clicking on the slides.
The format tab includes drawing tools and picture tools. The picture tool is a
context sensitive tab that appears on the Ribbon and allows the user to work with
inserted images, photos, Clip Art and pictures. It sets the brightness of images,
crops the picture, etc.
Navigation
Navigation through the slides can be accomplished using the Slide Navigation
menu on the left side of the screen. An outline of slides appear on the left side that
have been entered in the presentation. You need to click the outline tab to access
the outline of the presentation.
Self-Instructional
316 Material
Mini Toolbar Applications of Computer
in Educational Research
A new feature in Office 2007 is the Mini Toolbar. This is a floating toolbar that is
displayed when you select text or right click on the text. It displays the common
formatting tools, such as bold, italics, fonts, font size and font color. NOTES
Table
This option includes adding borders, rows and columns, formatting of individual
cell, etc., to a table. It helps in deciding the number of rows and columns that
would appear on the screen. The merge option combines the cells into a larger
one and the alignment option sets the alignment of the cell so that the text may fit
better.
Self-Instructional
Material 317
Applications of Computer
in Educational Research
NOTES
You can add or remove the commands from the list. Once you make changes
and save them, the quick access toolbar gets updated. This toolbar is also known
as a customizable toolbar. You can add or delete the toolbar from the menu.
Customize
Microsoft PowerPoint 2007 facilitates customizable options. For this, click on the
File menu. It generates the option called ‘PowerPoint Options’ as follows:
NOTES
The Popular option helps you to initialize the work environment, color
schemes and user name along with initials and accessing the Live Preview feature
which is useful for applying designs and changes.
The Proofing option provides auto correction settings and also helps in
finding errors through custom dictionaries.
The save option allows you to personalize the process of saving the
workbook.
The Advance feature option provides the options to edit, copy, paste, print,
display slide show and for general settings.
Self-Instructional
Material 319
Applications of Computer
in Educational Research
NOTES
The Customize option allows you to add or delete the toolbars in the quick
access toolbar. This option is very useful from the point of view of setting the
toolbars as per the user requirement.
You have already learnt the tools of application software which can be used for
documentation and making reports. In this section, you will study some general
guidelines related to document and report making.
Guidelines for Effective Documentation
Command over the medium: Even though one may have done an extremely
rigorous and significant research study, the fundamental test still remains as to how
the learning has been disseminated. Regardless of how effective the graphs and
figures are in showcasing the findings, the verbal description and explanation—in
terms of why it was done, how it was done, and what was the outcome, still
remain the acid test.
Thus, a correct and effective language of communication is critical in putting
ideas and objectives in the vernacular of the reader/decision-maker. The writer
may, thus, be advised to read professionally written reports and, if necessary,
seek assistance from those proficient in preparing business reports.
Phrasing protocol: There is a debate about whether or not one makes use
of personal pronoun while reporting. To understand this, one needs to revisit the
Self-Instructional
responsibility of the researcher, which is to present the findings of his/her study,
320 Material
with complete objectivity and precision. The use of personal pronoun such as ‘I Applications of Computer
in Educational Research
think…..’ or ‘in my opinion…..’ lends a subjectivity and personalization of
judgement. Thus, the tone of the reporting should be neutral. For example:
‘Given the nature of the forecasted growth and the opinion of the respondents,
NOTES
it is likely that the……’
Whenever the writer is reproducing the verbatim information from another
document or comment of an expert or published source, it must be in inverted
commas or italics and the author or source should be duly acknowledged.
The writer should avoid long sentences and break up the information in
clear chunks, so that the reader can process it with ease. Similar is the case in
structuring of the chapters or sections of the report that can be logically broken
down into smaller sections that are comprehensive and complete and yet maintain
a strong but logical link with the flow of reporting.
With the onset of the use of abbreviated communications in SMS and emails,
most people tend to use shortened form as ‘cd.’ for could and ‘u’ for you, etc.
Also the use of colloquial language and slangs must be avoided, as this is a formal
document and one must maintain the sanctity of the formal documentation required
in a research report.
Simplicity of approach: Along with grammatically and structurally correct
language, care must be taken to avoid technical jargon as far as possible. The
business manager, might have been a business student who had prepared a research
report in his academic pursuits but now understands simple common terms and
does not have the time or inclination to juggle the dictionary and the report together.
In case it is imperative to use certain terminology, then, as stated earlier, the definition
of these terms can be provided in the glossary of terms at the end of the report.
Sometimes the writer may prepare different research reports for the same
study to suit the need of diverse readers, for example, the business report needs to
be crisp and simple with definable and workable recommendations. On the other
hand, an academic report could discuss extensively the literature review section,
as well as the statistical analysis and interpretation.
Report formatting and presentation: In terms of paper quality, page
margins and font style and size, a professional standard should be maintained. The
font style must be uniform throughout the report. The topics, subtopics, headings
and subheadings must be construed in the same manner throughout the report.
Sometimes certain academic reports have a mandated format for presentation
which the writers need to follow, in which case there is no choice in presentation.
However, when this is not clear, it is advisable that the writer creates his/her
own formatting rules and saves it on a notepad so that they can be implemented in
a standardized and professional manner.
The researcher can provide data relief and variation by adequately
supplementing the text with graphs and figures. Pictorial representations are simple
Self-Instructional
Material 321
Applications of Computer to comprehend and also break the monotony and fatigue of reading. They should
in Educational Research
be used effectively whenever possible in the report.
Guidelines for Presenting Tabular Data
NOTES Most research studies involve some form of numerical data, and even though one
can discuss this in text, it is best represented in tabular form. The advantage of
doing this is that statistical tables present the data in a concise and numeral form,
which makes quantitative analysis and comparisons easier. Tables formulated could
be general tables following a statistical format for a particular kind of analysis.
These are best put in the appendix, as they are complex and detailed in nature.
The other kind is simple summary tables, which only contain limited information
and yet, are, essentially critical to the report text.
Table identification details: The table must have a title and an identification
number. The table title should be short and usually would not include any verbs or
articles. It only refers to the population or parameter being studied. The title should
be briefly yet clearly descriptive of the information provided. The numbering of
tables is usually in a series and generally one makes use of Arabic numbers to
identify them.
Data arrays: The arrangement of data in a table is usually done in an
ascending manner. This could either be in terms of time, (column-wise) or according
to sectors or categories (row-wise) or locations, e.g., north, south, east, west and
central. Sometimes, when the data is voluminous, it is recommended that one
goes alphabetically, e.g., country or state data. Sometimes there may be
subcategories to the main categories, for example, under the total sales data—a
column-wise component of the revenue statement—there could be subcategories
of department store, chemists and druggists, mass merchandisers and others.
Measurement unit: The unit in which the parameter or information is
presented should be clearly mentioned.
Spaces, Leaders and Rulings (SLR): For limited data, the table need
not be divided using grid lines or rulings. Simple white spaces add to the clarity of
information presented and processed. In case the number of parameters are too
many and the data seems to be bulky to be simply separated by space, it is advisable
to use vertical ruling. Horizontal lines are drawn to separate the headings from the
main data. When there are a number of subheadings as in the sales data example,
one may consider using leaders (…….) to assist the eye movement in absorbing
and processing the information.
Total sales
Mass market………
Department store………
Drug stores………
Others (including paan beedi outlets)………
Self-Instructional
322 Material
Assumptions, details and comments: Any clarification or assumption Applications of Computer
in Educational Research
made, or a special definition required to understand the data, or formula used to
arrive at a particular figure, e.g., total market sale or total market size can be given
after the main tabled data in the form of footnotes.
NOTES
Data sources: In case the information documented and tabled is secondary
in nature, complete reference of the source must be cited after the footnote, if any.
Special mention: In case some figure or information is significant and the
reader should pay special attention to it, the number or figure can be bold or can
be highlighted to increase focus.
Guidelines for Visual Representations: Graphs
Similar to the summarized and succinct data in the form of tables, the data can also
be presented through visual representations in the form of graphs. The visual
representation of the findings in the form of lines or boxes and bars relative to a
number line is easy to comprehend and interpret. There are some standard rules
and procedures available to the researcher for this; also there are computer
programs like MS Excel and SPSS, where the numbered data can be converted
with ease into graphical form.
Line and curve graphs: Usually, when the objective is to demonstrate
trends and some sort of pattern in the data, a line chart is the best option available
to the researcher as the line is able to clearly portray any change in pattern during
a particular time period. On the same chart, it is also possible to show patterns of
growth of different sectors or industries in the same time period or to compare the
change in the studied variable across different organizations or brands in the same
industry. Certain points to be kept in mind while formulating line charts include:
The time units or the causal variable being studied are to be put on the X-
axis, or the horizontal axis.
If the intention is to compare different series on the same chart, the lines
should be of different colours or forms.
Too many lines are not advisable on the same chart as then the data becomes
too cluttered; an ideal number would be five or less than five lines on the
chart.
The researcher also must take care to formulate the zero baseline in the
chart as otherwise, the data would seem to be misleading.
Area or stratum charts: Area charts are like the line charts, usually used
to demonstrate changes in a pattern over a period of time. However, here there
are multiple lines that are essentially components of the original composite data.
What is done is that the change in each of the components is individually shown on
the same chart and each of them is stacked one on top of the other. The areas
between the various lines indicate the scale or volume of the relevant factors/
categories.
Self-Instructional
Material 323
Applications of Computer Pie charts: Another way of demonstrating the area or stratum or sectional
in Educational Research
representation is through the pie charts. The critical difference between a line and
pie chart is that the pie chart cannot show changes over time. It simply shows the
cross-section of a single time period. The sections or slices of the pie indicate the
NOTES ratio of that section to the total area of the parameter being displayed. There are
certain rules that the researcher should keep in mind while creating pie charts.
The complete data must be shown as a 100 per cent area of the subject
being graphed.
It is a good idea to have the percentages displayed within or above the pie
rather than in the legend as then it is easier to understand the magnitude of
the section in comparison to the total.
Bar charts and histograms: A very useful representation of quantum or
magnitude of different objects on the same parameter are bar diagrams. The
comparative position of objects becomes very clear. The usual practice is to
formulate vertical bars; however, it is possible to use horizontal bars as well if
none of the variable is time related. Horizontal bars are especially useful when one
is showing both positive and negative patterns on the same graph. These are called
bilateral bar charts and are especially useful to highlight the objects or sectors
showing a varied pattern on the studied parameter. It is possible to generate bar
graphs with relative ease with computer programs today and the distance between
the bars can be extremely precise as compared to those created by hand.
Another variation of the bar chart is the histogram here the bars are vertical
and the height of each bar reflects the relative or cumulative frequency of that
particular variable.
Pictogram: A pictogram shows graphical representation of data. Pictograms
are most often used in popular and general read such as in magazines and
newspapers, as they are eye-catching and easy to comprehend by one and all.
They are not a very accurate or scientific representation of the actual data and,
thus, should be used with caution in an academic or technical report.
Geographic representation: Geographic or regional maps related to
countries, states, districts, territories can be used as a base to show occurrence of
the studied variable in various regions or to show comparative analysis about
major brands or industries or minerals. In case of comparative data, the researcher
must provide the legend in the displayed map, for example any map of the location
may be given.
For the moment, we will concentrate on the second option, i.e., Type in
data. Select this option and click Ok. By default, the Data Editor view is initially
selected.
SPSS Data Editor
The SPSS Data Editor Window has two views: Data View and Variable View.
Variable View is used to define variables that will store the data. Data View contains
the actual data.
The first step is to open the ‘Variable View’ window of the Data Editor and
define variables. Let us consider an example where Employee Data of an
organization needs to be saved and analysed. The objective is to create a small
data file for employees that consist of six variables as given in Table below.
EmpID Numeric
EmpName String
Age Numeric
Income Numeric
Self-Instructional
326 Material
There are different types of variables in SPSS, the default one being numeric. Applications of Computer
in Educational Research
To change variable type, in Variable View click on the variable in the column Type.
A window similar to one below will open. Create all the variables and select
appropriate Type as given in the table above.
NOTES
Note: While defining variable names empty spaces are not allowed.
E.g., Marital Status – Not allowed
MaritalStatus or Marital_Status – Correct
The third column in Variable view is Width, which specifies the number of
characters allowed to be entered in the column. By default the width is 8 characters
and can be modified depending upon the data being entered.
The fourth column is Decimals, which represents the number of decimal
places. For numeric data type the default value is 0. Say, for example, EmpID
does not require decimal places, therefore, it can be set to 0.
The fifth column is Label, which describes the variable.
The sixth column is Values. For example, Gender contains two categories
(Female = 1 and Male = 2). In Data View, the gender will be entered as either 1
or 2. But what 1 or 2 represents is given in the Values as 1 represents Female and
2 Male.
The seventh column is Missing. Often while collecting data, you will have
missing values within your data. This column is used in cases where no data is
provided by a respondent. A missing value is chosen as an impossible value for
that column. For example, the missing value for age can be entered as 1000 or -
100 which are impossible entries for age. The objective of giving a missing value is
to exclude that record while analysing the data.
Self-Instructional
Material 327
Applications of Computer The eighth column is Columns. It represents the width of the column. Default
in Educational Research
value is 8 and can be changed.
The ninth column is Align, which aligns the data at the left, centre or right of
NOTES cell.
The last column is Measure. It can take values of Nominal, Ordinal or
Scale.
The table below shows the different types of measurement, with examples:
Enter some data for the variables created in the Variable View. The Data
View grid will look something like shown below:
Self-Instructional
328 Material
Applications of Computer
in Educational Research
NOTES
Recoding Variables
Recode is a very important feature in SPSS, which is used to convert continuous
data into discrete or category data. One can recode values within the existing
variable into a new variable.
Note: If you recode the values into the existing variable, the old values are lost. So it is
recommended to recode a variable into a new variable wherever possible, so that your
original values are retained.
Recode is available under Transform menu. There are three ways to recode
the data.
1. Recode into same variables
2. Recode into new variables
3. Automatic recode
Now suppose, the variable income is to be categorized into three income
categories based upon the below logic.
< =10000 – 1 (Low income)
>10000 - <=30000 - 2 (Middle income)
> 30000 as 3 (High income)
Go to Transform-> Recode into new variable. The variable income will be
recoded into a new variable (IncomeRe) labeled as Income Redefined which is
the Output Variable.
Self-Instructional
Material 329
Applications of Computer Click on the button Old and New Values. A window will open divided into
in Educational Research
two parts. Left side will be Old Value and right side shows New Value.
Since the first category is 10000, the Old Value option to be selected will
be Range, Lowest through value: 10,000. New Value is 1.
NOTES
The second category is a range >10000 and 30,000, the Old Value option
to be selected is a Range, i.e., 10,000 through 30,000. New Value is 2.
The third category is > 3000, the Old value option to be selected is Range,
value through Highest: 30,000. New Value is 3.
A snapshot of the recode screen is given below for reference. Click on
Continue and Ok.
A new variable IncomeRe will be created based upon the income variable.
Next, we need to label what are 1, 2 and 3 values. Go to Variable View and give
the labels for the new variable IncomeRe.
There are a number of specific software programs like E Views for business
forecasting and LISREL Linear Structural Relations) for structural equation
modelling. However, for most purposes, SPSS is the most widely used software.
Self-Instructional
330 Material
Applications of Computer
14.9 ANSWERS TO CHECK YOUR PROGRESS in Educational Research
QUESTIONS
1. The major components of MS Office include Excel, Word, One Note and NOTES
PowerPoint.
2. Ribbon is a panel in MS Office which houses a fixed arrangement of few
icons and command buttons.
3. The options of the default commands like Save, Undo and Redo are available
on the Quick Access Tool in MS Word.
4. The alternative key command for opening an existing document on MS
Word is CTRL+O on the keyword.
5. The Proofing feature in MS Excel permits you to personalize the options for
correcting words and formats of your text.
6. Th keyboard function CTRL+F4 in MS Excel closes a workbook file.
7. The slide show function in MS PowerPoint helps in setting up the start slide
show and to record narration. It also has the option to monitor screen
resolution by providing separate views of the slide show.
8. The table title short be short and usually wold not include any verb and
articles. It only refers to the population or parameter being studied.
9. Area charts are like the line charts, usually used to demonstrate changes in
a pattern over a period of time.
10. Linear models, generalized linear models, multivariate methods, categorical
data analysis, and all the standard techniques for descriptive and confirmatory
statistical analysis are possible with SAS.
14.10 SUMMARY
Self-Instructional
332 Material
Similar to the summarized and succinct data in the form of tables, the data Applications of Computer
in Educational Research
can also be presented through visual representations in the form of graphs.
The visual representation of the findings in the form of lines or boxes and
bars relative to a number line is easy to comprehend and interpret. There
are some standard rules and procedures available to the researcher for this; NOTES
also there are computer programs like MS Excel and SPSS, where the
numbered data can be converted with ease into graphical form.
Researchers have to their advantage a wide array of statistical programmes
to assist them in both data management and data analysis. In this section we
will briefly discuss only the most frequently used packages.
Self-Instructional
Material 333
Applications of Computer
in Educational Research 14.13 FURTHER READINGS
Best, J.W. and J.V. Kahn. 1989. Research in Education. New Delhi: Prentice
NOTES Hall.
Buch, M. B. 1974. A Survey of Research in Education. Baroda: CASE, M. S.
University.
Fox, D. J. 1969. The Research Process in Education. New York: Rhinehart and
Winston, Inc.
Garrett. H. E. 1988. Statistics in Psychology and Education. Bombay: Vikils,
Feiffer & Semen’s Ltd.
Guilford, J.P. and B. Fruchter. 1974. Fundamental Statistics in Psychology &
Education. New York: McGraw Hill.
Self-Instructional
334 Material
High levels of non-sampling errors compromise the accuracy and reliability of survey research. These errors, arising from factors like survey design flaws, respondent misinterpretation, or data processing mistakes, can skew results significantly. To minimize non-sampling errors, researchers should ensure precise questionnaire design, rigorous training for interviewers, pre-testing surveys, and implementing checks during data processing. Attention to survey administration and methodology will help reduce these errors, leading to more reliable and valid survey outcomes .
The concept of statistical power, which is the probability of correctly rejecting a false null hypothesis (1 - β), plays a crucial role in determining the sample size for hypothesis testing. High statistical power is desirable as it increases the likelihood of detecting true effects in the population. To achieve a desired level of power, researchers often need to increase sample sizes, especially when expecting small effect sizes or when dealing with high variability in data. Thus, power analysis is typically conducted pre-study to estimate the appropriate sample size needed to achieve reliable test results .
When deciding between probability and non-probability sampling, factors to consider include the study’s objectives, the type of study, available resources, and the desired generalizability of findings. Probability sampling is ideal when researchers aim for generalizable results with known precision, requiring a larger budget and timeframe due to its complexity and potential for requiring larger samples. Non-probability sampling suits exploratory research where quick, cost-effective insights are needed, but with limitations in the generalizability of findings and potential biases .
Probability sampling enhances the reliability of research findings by ensuring that every member of the population has a known, non-zero chance of being included in the sample. This method reduces bias and allows for the generalization of results to the entire population. It relies on randomness, which mitigates selection bias and provides a basis for the application of probability theory to make statistical inferences about the population. In contrast, non-probability sampling lacks this randomness, leading to potential biases and less reliable generalizations .
Ethnography distinguishes itself by focusing on the comprehensive exploration of cultural phenomena within a community from an insider's viewpoint (emic perspective). Unlike other qualitative approaches, ethnography involves prolonged engagement and immersive observation within the community being studied. This method aims to describe cultural meanings and practices in the context of everyday life and is highly descriptive and interpretive. Other qualitative methods, such as case studies or phenomenology, may not involve as extensive immersion and often have a narrower focus or are more structured .
Common sources of sampling errors include bias in selection processes, errors in estimation, and variability due to the subset nature of samples. Sampling errors occur when a sample does not accurately represent the population, often due to a small sample size or unrepresentative sampling methods. To mitigate these errors, researchers can increase sample sizes, use probability sampling methods to ensure randomness, and apply stratified sampling to ensure subgroup representation. These strategies help reduce the discrepancies between the sample and the population estimates .
The Kruskal-Wallis test acts as a non-parametric alternative to one-way ANOVA, useful when the assumption of normal distribution in the populations is not met. It is particularly applicable for ordinal data or when sample sizes are small, where assuming normality would be invalid. The Kruskal-Wallis test compares the medians of k independent samples, determining if they are from identical populations without assuming a normal distribution. This makes it suitable for data that violates ANOVA's assumptions, ensuring robust and applicable statistical conclusions despite dataset limitations .
Non-probability sampling methods include incidental or accidental sampling, judgement sampling, purposive sampling, quota sampling, and snowball sampling. These methods are primarily based on subjective judgment rather than random selection. Consequently, they are prone to various biases: selection bias due to non-random sampling, sampling bias because of over-reliance on the researcher's judgment, and response bias due to self-selection. These biases limit the generalizability of findings to a larger population and increase the risk of drawing inaccurate conclusions .
Systematic sampling is more advantageous than simple random sampling when there is a defined population list and when efficiency in sample selection is crucial. It is practical for large populations to alleviate the cumbersome nature of generating random numbers. Additionally, systematic sampling can ensure a spread across the population, minimizing clustering effects. However, this method assumes no hidden pattern in the population that aligns with the sampling interval, which could introduce bias. Thus, it is most useful when a listed population is naturally void of hidden periodicities .
Setting a level of significance, represented by alpha (α), is crucial in hypothesis testing because it defines the threshold for rejecting a null hypothesis. It signifies the probability of committing a Type I error, which occurs when a true null hypothesis is wrongly rejected. A commonly used α value is 0.05, implying a 5% risk of making such an error, and establishes a confidence level of 95% for the decision. The level of significance directly influences the critical value for test statistics, thus affecting the conclusions drawn from a statistical test .