1.
Introduction to Sampling
Sampling is the process of selecting a subset (sample) from a larger population to estimate
characteristics of the whole population.
Why Sampling?
Saves time and cost
Practical for large populations
Enables statistical inference
Example (Agricultural Context):
Instead of surveying all farmers in a region, a sample is selected to estimate average crop yield.
2. Key Concepts
Population: Entire group of interest (e.g., all maize farmers in a district)
Sample: Subset of the population
Sampling Frame: List of population elements
Sampling Unit: Individual element (e.g., a farm)
3. Types of Sampling Techniques
A. Probability Sampling
Each unit has a known, non-zero chance of selection.
3.1 Simple Random Sampling
Definition:
Every member has an equal chance of being selected.
Methods:
Lottery method
Random number tables/software
Advantages:
Easy to understand
Minimizes bias
Disadvantages:
Requires complete list of population
Can be costly for large populations
Agricultural Application:
Selecting farmers randomly to estimate average fertilizer use.
3.2 Stratified Sampling
Definition:
Population divided into homogeneous subgroups (strata), then sampled from each.
Steps:
1. Divide population (e.g., by farm size: small, medium, large)
2. Sample from each group
Advantages:
More precise estimates
Ensures representation
Disadvantages:
Requires knowledge of population structure
Agricultural Application:
Comparing productivity across irrigation vs rain-fed farms.
3.3 Systematic Sampling
Definition:
Selecting every k-th unit from a list.
Formula:
Example:
If N = 1000 farms, n = 100 → select every 10th farm.
Advantages:
Simple and quick
Even coverage
Disadvantages:
Risk of periodic bias
Agricultural Application:
Surveying every 5th household in a farming village.
3.4 Cluster Sampling
Definition:
Population divided into clusters (e.g., villages), then randomly select clusters.
Advantages:
Cost-effective
Useful for geographically dispersed populations
Disadvantages:
Less precise than stratified sampling
Agricultural Application:
Selecting 10 villages and surveying all farmers within them.
4. Non-Probability Sampling
Not based on random selection.
4.1 Convenience Sampling
Definition:
Selecting easily available units.
Example:
Interviewing farmers at a local market.
Limitations:
High bias
Not representative
4.2 Judgmental (Purposive) Sampling
Definition:
Researcher selects units based on expertise.
Example:
Selecting experienced farmers for expert opinion.
4.3 Quota Sampling
Definition:
Ensuring certain characteristics are represented.
Example:
50% male farmers, 50% female farmers.
5. Sampling Errors
5.1 Sampling Error
Difference between sample estimate and population value.
5.2 Non-Sampling Errors
Measurement errors
Data entry mistakes
Non-response bias
6. Sample Size Determination
Factors affecting sample size:
Population size
Variability
Desired accuracy
Confidence level
7. Choosing the Right Sampling Technique
Situation Recommended Method
Homogeneous population Simple Random
Diverse population Stratified
Large geographic area Cluster
Ordered list available Systematic
Limited resources Convenience
8. Applications in Agricultural Economics
8.1 Crop Yield Estimation
Use stratified sampling across regions
8.2 Farm Income Studies
Random sampling for unbiased estimates
8.3 Market Surveys
Systematic sampling of traders
8.4 Policy Impact Analysis
Cluster sampling of villages
8.5 Risk & Uncertainty Studies
Sampling farmers to analyze weather impacts
9. Practical Example
Problem: Estimate average maize yield in a district.
Solution Approach:
1. Divide farms into strata (small, medium, large)
2. Randomly sample from each group
3. Compute weighted average
10. Advantages of Proper Sampling
Reduces cost
Improves decision-making
Enables policy formulation
Supports agricultural planning
11. Limitations of Sampling
Sampling bias
Requires expertise
May not capture rare events
12. Summary
Sampling is essential in agricultural research
Probability sampling ensures accuracy
Choice of method depends on research goals
Proper design improves reliability of results