0% found this document useful (0 votes)
180 views12 pages

Biometry: Sokal & Rohlf 4th Edition PDF

This document provides information about the book "Biometry: The Principles and Practice of Statistics in Biological Research" by Robert R. Sokal and F. J. Rohlf. It includes the publication details, author profiles, number of citations, and related projects of one of the authors. The document then provides the front matter content of the book, including its dedication and table of contents.

Uploaded by

Alisa Rivera
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd
0% found this document useful (0 votes)
180 views12 pages

Biometry: Sokal & Rohlf 4th Edition PDF

This document provides information about the book "Biometry: The Principles and Practice of Statistics in Biological Research" by Robert R. Sokal and F. J. Rohlf. It includes the publication details, author profiles, number of citations, and related projects of one of the authors. The document then provides the front matter content of the book, including its dedication and table of contents.

Uploaded by

Alisa Rivera
Copyright
© All Rights Reserved
We take content rights seriously. If you suspect this is your content, claim it here.
Available Formats
Download as PDF, TXT or read online on Scribd

See discussions, stats, and author profiles for this publication at: [Link]

net/publication/44554870

Biometry : the principles and practice of statistics in biological


research / Robert R. Sokal and F. James Rohlf

Book · April 2013


Source: OAI

CITATIONS READS

51 58,888

2 authors, including:

F. James Rohlf
State University of New York
255 PUBLICATIONS 61,501 CITATIONS

SEE PROFILE

Some of the authors of this publication are also working on these related projects:

book reviews View project

Numerical taxonomy View project

All content following this page was uploaded by F. James Rohlf on 03 March 2014.

The user has requested enhancement of the downloaded file.


Biometry FOURTH EDITION

[Link] i 8/5/11 1:07 PM


[Link] ii 8/5/11 1:07 PM
Biometry
The Principles and Practice of Statistics
in Biological Research
FOURTH EDITION

Robert R. Sokal and F. James Rohlf


Stony Brook University

W.H. Freeman and Company


New York

[Link] iii 8/5/11 1:07 PM


Publisher: Peter Marshall
Acquisitions Editor: Jerry Correa
Marketing Manager: Debbie Clare
Project Editor: Marni Rolfes
Art Director: Diana Blume
Project management, Illustrations, and Composition: MPS Limited, a Macmillan
Company
Production Coordinator: Lawrence Guerra
Printing and Binding: RR Donnelley

Library of Congress Control Number: 2010939805

ISBN-13: 978-0-7167-8604-4
ISBN-10: 0-7167-8604-4

©2012, 1995, 1981, 1969 by W. H. Freeman and Company


All rights reserved

Printed in the United States of America

First printing

W.H. Freeman and Company


41 Madison Avenue
New York, NY 10010
Houndmills, Basingstoke RG21 6XS, England
[Link]

[Link] iv 8/11/11 10:19 AM


To our wives
Julie and Janice

[Link] v 8/5/11 1:07 PM


[Link] vi 8/5/11 1:07 PM
Contents

Preface xiii
Notes on the Fourth Edition xvii

1 Introduction 1
1.1 Some Definitions 1
1.2 The Development of Biometry 3
1.3 The Statistical Frame of Mind 5

2 Data in Biology 9
2.1 Samples and Populations 9
2.2 Variables in Biology 11
2.3 Accuracy and Precision of Data 13
2.4 Derived Variables 16
2.5 Frequency Distributions 19

3 Computers and Data Analysis 33


3.1 Computers 33
3.2 Software 35
3.3 Efficiency and Economy in Data Processing 37

4 Descriptive Statistics 39
4.1 The Arithmetic Mean 40
4.2 Other Means 44
4.3 The Median 45
4.4 The Mode 47
4.5 Sample Statistics and Parameters 49
4.6 The Range 49
4.7 The Standard Deviation 51
4.8 Coding Data Before Computation 54
4.9 The Coefficient of Variation 55

vii

[Link] vii 8/5/11 1:07 PM


viii Contents

5 Introduction to Probability Distributions:


Binomial and Poisson 59
5.1 Probability, Random Sampling, and Hypothesis Testing 60
5.2 The Binomial Distribution 68
5.3 The Poisson Distribution 78
5.4 Other Discrete Probability Distributions 87

6 The Normal Probability Distribution 93


6.1 Frequency Distributions of Continuous Variables 93
6.2 Properties of the Normal Distribution 95
6.3 A Model for the Normal Distribution 100
6.4 Applications of the Normal Distribution 102
6.5 Fitting a Normal Distribution to Observed Data 104
6.6 Skewness and Kurtosis 106
6.7 Graphic Methods 108
6.8 Other Continuous Distributions 117

7 Hypothesis Testing and Interval Estimation 119


7.1 Introduction to Hypothesis Testing: Randomization Approaches 120
7.2 Distribution and Variance of Means 131
7.3 Distribution and Variance of Other Statistics 137
7.4 The t-Distribution 140
7.5 More on Hypothesis Testing: Normally Distributed Data 142
7.6 Power of a Test 146
7.7 Tests of Simple Hypotheses Using the Normal and t-Distributions 148
7.8 The Chi-Square Distribution 154
7.9 Testing the Hypothesis H0: s2 5 s02 156
7.10 Introduction to Interval Estimation (Confidence Limits) 157
7.11 Confidence Limits Using Sample Standard Deviations 162
7.12 Confidence Limits for Variances 167
7.13 The Jackknife and the Bootstrap 168

8 Introduction to Analysis of Variance 177


8.1 Variances of Samples and Their Means 178
8.2 The F-Distribution 182
8.3 The Hypothesis H0: s12 5 s22 187
8.4 Heterogeneity Among Sample Means 190
8.5 Partitioning the Total Sum of Squares
and Degrees of Freedom 197

[Link] viii 8/5/11 1:07 PM


Contents ix

8.6 Model I Anova 200


8.7 Model II Anova 203

9 Single-Classification Analysis of Variance 207


9.1 Computational Formulas 208
9.2 General Case: Unequal and Equal n 208
9.3 Special Case: Two Groups 220
9.4 Comparisons Among Means in a Model I Anova: Essential
Background 228
9.5 Comparisons Among Means: Special Methods 246

10 Nested Analysis of Variance 277


10.1 Nested Anova: Design 277
10.2 Nested Anova: Computation 280
10.3 Nested Anovas with Unequal Sample Sizes 301

11 Two-Way and Multiway Analysis of Variance 319


11.1 Two-Way Anova: Design 319
11.2 Two-Way Anova with Equal Replication: Computation 321
11.3 Two-Way Anova: Hypothesis Testing 331
11.4 Two-Way Anova Without Replication 340
11.5 Paired Comparisons 349
11.6 The Factorial Design 354
11.7 A Three-Way Factorial Design 355
11.8 Higher-Order Factorial Anovas 365
11.9 Other Designs 370
11.10 Anova by Computer 372

12 Statistical Power and Sample Size in the


Analysis of Variance 379
12.1 Effect Size 379
12.2 Noncentral t- and F-Distributions and Confidence Limits
for Effect Sizes 382
12.3 Power in an Anova 390
12.4 Sample Size in an Anova 391
12.5 Minimum Detectable Difference 395
12.6 Post Hoc Power Analysis 396
12.7 Optimal Allocation of Resources in a Nested Design 397
12.8 Randomized Blocks and Other Two-Way and Multiway Designs 406

[Link] ix 8/5/11 1:07 PM


x Contents

13 Assumptions of Analysis of Variance 409


13.1 A Fundamental Assumption 410
13.2 Independence 410
13.3 Homogeneity of Variances 413
13.4 Normality 422
13.5 Transformations 426
13.6 The Logarithmic Transformation 427
13.7 The Square Root Transformation 433
13.8 The Box–Cox Transformation 435
13.9 The Arcsine Transformation 438
13.10 Nonparametric Methods in Lieu of Single-Classification
Anova 440
13.11 Nonparametric Methods in Lieu of Two-Way Anova 460

14 Linear Regression 471


14.1 Introduction to Regression 472
14.2 Models in Regression 475
14.3 The Linear Regression Equation 477
14.4 Hypothesis Testing in Regression 485
14.5 More Than One Value of Y for Each Value of X 495
14.6 The Uses of Regression 506
14.7 Estimating X From Y 511
14.8 Comparing Two Regression Lines 513
14.9 Linear Comparisons in Anovas 515
14.10 Examining Residuals and Transformations in Regression 524
14.11 Nonparametric Tests for Regression 532
14.12 Model II Regression 535
14.13 Effect Size, Power, and Sample Size in Regression 544

15 Correlation 551
15.1 Correlation Versus Regression 551
15.2 The Product–Moment Correlation Coefficient 554
15.3 Computing the Product–Moment Correlation Coefficient 562
15.4 The Variance of Sums and Differences 565
15.5 Hypothesis Tests for Correlations 567
15.6 Applications of Correlation 577
15.7 Nonparametric Tests for Association 580
15.8 Major Axes and Confidence Regions 588
15.9 Effect Size, Power, and Sample Size 592

[Link] x 8/12/11 6:49 PM


Contents xi

16 Multiple and Curvilinear Regression 603


16.1 Multiple Regression: Computation 604
16.2 Multiple Regression: Hypothesis Tests 614
16.3 Path Analysis and Structural Equation Modeling 625
16.4 Partial and Multiple Correlation 644
16.5 Selection of Independent Variables 649
16.6 Computation of Multiple Regression by Matrix Methods 656
16.7 Solving Anovas as Regression Problems:
General Linear Models 659
16.8 Analysis of Covariance (Ancova) 665
16.9 Curvilinear Regression 671
16.10 Effect Size, Power, and Sample Size in Multiple Regression 685
16.11 Advanced Topics in Regression and Correlation 694

17 Analysis of Frequencies 703


17.1 Introduction to Tests for Goodness of Fit 704
17.2 Single-Classification Tests for Goodness of Fit 714
17.3 Replicated Tests of Goodness of Fit 730
17.4 Tests of Independence: Two-Way Tables 739
17.5 Analysis of Three-Way Tables 758
17.6 Analysis of Proportions 773
17.7 Randomized Blocks for Frequency Data 793
17.8 Effect Sizes, Power, and Sample Sizes 801

18 Meta-Analysis and Miscellaneous Methods 817


18.1 Synthesis of Prior Research Results: Meta-Analysis 817
18.2 Tests for Randomness of Nominal Data: Runs Tests 841
18.3 Isotonic Regression 847
18.4 Application of Randomization Tests to Unconventional Statistics 850
18.5 The Mantel Test of Association Between Two Distance Matrices 852
18.6 The Future of Biometry: Data Analysis 859

Appendices
A. Mathematical Proofs 869
B. Introduction to Matrices 885
Bibliography 891
Author Index 909
Subject Index 915

[Link]
View publication stats xi 8/5/11 1:07 PM

Common questions

Powered by AI

The normal distribution is applied in situations involving continuous variables and is characterized by its bell-shaped curve and specific properties like mean and standard deviation . It is highly useful in biological research for assessing the distribution of continuous biological data. In contrast, the binomial and Poisson distributions are used for discrete variables. The binomial distribution applies when there are fixed numbers of trials with two possible outcomes per trial, such as success and failure . The Poisson distribution is used for events occurring independently over a fixed interval, such as the number of mutations in a given length of DNA .

Meta-analysis advances biological research by aggregating results from multiple studies, thereby increasing statistical power and improving effect size estimation . It helps in identifying patterns or inconsistencies across studies, providing a more comprehensive view of the topic under examination. By synthesizing findings, meta-analysis enhances generalizability and can reveal overarching conclusions that might not be apparent from singular studies. This is particularly valuable in biology, where individual study outcomes may vary due to species differences, environmental factors, or methodological variability .

Transformations such as logarithmic, square root, and Box-Cox transformations are used to address violations in the assumptions of ANOVA, particularly regarding normality and homogeneity of variances . These transformations help stabilize variance and make the data conform more closely to a normal distribution, which is an assumption of ANOVA . Applying transformations is necessary to ensure the validity of the results, as ANOVA requires these assumptions to be met for its conclusions to be statistically reliable.

Nonparametric methods are used when parametric test assumptions, such as normality and homoscedasticity, are violated . These methods do not rely on specific distributional assumptions and are therefore more flexible, making them suitable for analyzing data that deviate from parametric prerequisites. In biological research, nonparametric methods can provide valid inferences by ranking data or using median-based statistics, which are less sensitive to outliers and skewed distributions . They are crucial for maintaining the rigor of statistical conclusions in studies where normality and variance homogeneity cannot be assured.

Key assumptions underlying the linear regression model include linearity, independence, homoscedasticity (equal variance), and normally distributed residuals . Violations of these assumptions can lead to biased and inefficient estimates. For example, non-linearity might result in poor model fit, while non-independence can inflate Type I error rates. Homoscedasticity is vital for consistent and reliable estimates; its violation can lead to underestimated variances . Ensuring these assumptions are met or addressing violations through techniques like transformations or robust regression methods is crucial for maintaining the integrity of the regression analysis results.

Homogeneity of variances, or the assumption that different groups in the analysis have the same variance, is critical for the validity of ANOVA results . When this assumption is violated, it can lead to incorrect conclusions, as the F-test used in ANOVA is sensitive to differences in group variances . Variance homogeneity ensures that the comparison among group means is not skewed by unequal variability, thereby stabilizing the Type I error rate.

The Jackknife and Bootstrap methods are resampling techniques used in estimating confidence intervals, particularly when traditional assumptions do not hold . The Jackknife involves systematically omitting individual observations from the dataset to assess the variability of a statistical estimate, while the Bootstrap involves repeatedly sampling with replacement to create a distribution of the estimate . These methods offer the advantage of not requiring normality or equal variance assumptions, making them highly versatile and robust in a variety of biological research scenarios, especially for smaller or non-normal samples.

Confidence intervals for sample standard deviations are constructed using approaches such as the Chi-square distribution, which provides bounds within which the true population standard deviation is likely to lie . This statistical measure is crucial in biometry as it provides a range of values that have a high probability of containing the population parameter, thereby offering a way to assess the precision of the sample estimate and the variability associated with it . Understanding the uncertainty around the estimated standard deviation supports better decision-making and reliability in biological research.

The power of a statistical test, defined as the probability of correctly rejecting a false null hypothesis, is crucial in determining the test's sensitivity . High power reduces the risk of Type II errors and increases the chance of detecting true effects, thus affecting outcomes in biological research by ensuring meaningful differences or effects are not overlooked . A test with low power may lead researchers to miss important findings due to insufficient sensitivity, which could result in incorrect conclusions being drawn from an experiment.

In biological research, data precision refers to the consistency and repeatability of measurements, while statistical accuracy denotes closeness to the true value . High precision does not guarantee accuracy, as consistent results can still be systematically biased. Accuracy without precision, on the other hand, implies variability that can obscure true trends. Balancing both is critical; precise measurements must also be accurate to ensure reliable and valid conclusions. This relationship underscores the importance of robust experimental design and calibration in biological studies to mitigate errors and improve the trustworthiness of the findings .

You might also like