PubMed HealthSearch

PubMed · 9221444

[Estimation methods in a sampling survey].

Abstract

The objective of this paper is to present a guide for statistical analysis of a sampling survey. Advantages and disadvantages of classical sampling designs are first discussed. A sampling survey was designed to give more precise estimates of parameters which characterize a targeted well-defined population. Simple sampling design often is impossible in large population and rarely is the optimal solution. According to practical and economic constraints, and to available sampling frames, other strategies (stratification, unequal probabilities, several selection steps) may be necessary or more efficient. Fundamental tools to compute estimators and their confidence intervals are presented. The choice of the sampling method determine for each population unit the probability to include it in the final sample. These inclusion probabilities must be known to formulate estimators, ideally unbiased and with low variance. Difficulties arise from calculating the variance, especially for estimators which are not linear functions of the characteristics of interest. In that case, estimation procedures include Taylor linearization. It is always possible to find at least one linear estimator for a total. The total is the key-parameter in sampling theory, most parameters (such as ratios, means, percentages...), being function of unknown totals. Numerical examples are given.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

J Warszawski, J Lellouch. 1997. [Estimation methods in a sampling survey].. https://pubmed.ncbi.nlm.nih.gov/9221444/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Sample size determination for multiple comparison studies treating confidence interval width as random.

Methods for optimal sample size determination are developed using four popular multiple comparison procedures (Scheffe's, Bonferroni's, Tukey's and Dunnett's procedures), where random samples of the same size n are to be selected from k (>/=2) normal populations with common variance sigma2, and where primary interest concerns inferences about a family of L linear contrasts among the k population means. For a simultaneous coverage probability of (1-alpha), the optimal sample size is defined to be the smallest integer value n*m such that, simultaneously for all L confidence intervals, the width of the lth confidence interval will be no greater than tolerance 2deltal (l=1,2,...,L) with tolerance probability at least (1-gamma), treating the pooled sample variance S2p as a random variable. Using Scheffe's procedure as an illustration, comparisons are made to usual sample size methods that incorrectly ignore the stochastic nature of S2p. The latter approach can lead to serious underestimation of required sample sizes and hence to unacceptably low values of the actually tolerance probability (1-gamma'). Our approach guarantees a lower bound of [1-(alpha+gamma)] for the probability that the L confidence intervals will both cover the parametric functions of interest and also be sufficiently narrow. Recommendations are provided regarding the choices among the four multiple comparison procedures for sample size determination and inference-making.

Confidence Intervals