PubMed HealthSearch

PubMed · 9333340

Regression estimator in ranked set sampling.

Abstract

Ranked set sampling (RSS) utilizes inexpensive auxiliary information about the ranking of the units in a sample to provide a more precise estimator of the population mean of the variable of interest Y, which is either difficult or expensive to measure. However, the ranking may not be perfect in most situations. In this paper, we assume that the ranking is done on the basis of a concomitant variable X. Regression-type RSS estimators of the population mean of Y will be proposed by utilizing this concomitant variable X in both the ranking process of the units and the estimation process when the population mean of X is known. When X has unknown mean, double sampling will be used to obtain an estimate for the population mean of X. It is found that when X and Y jointly follow a bivariate normal distribution, our proposed RSS regression estimator is more efficient than RSS and simple random sampling (SRS) naive estimators unless the correlation between X and Y is low (/rho/ < 0.4). Moreover, it is always superior to the regression estimator under SRS for all rho. When normality does not hold, this approach could still perform reasonably well as long as the shape of the distribution of the concomitant variable X is only slightly departed from symmetry. For heavily skewed distributions, a remedial measure will be suggested. An example of estimating the mean plutonium concentration in surface soil on the Nevada Test Site, Nevada, U.S.A., will be considered.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

P L Yu, K Lam. 1997. Regression estimator in ranked set sampling.. https://pubmed.ncbi.nlm.nih.gov/9333340/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

On the calculation of a robust S-estimator of a covariance matrix.

An S-estimator of multivariate location and scale minimizes the determinant of the covariance matrix, subject to a constraint on the magnitudes of the corresponding Mahalanobis distances. The relationship between S-estimators and w-estimators of multivariate location and scale can be used to calculate robust estimates of covariance matrices. Elemental subsets of observations are generated to derive initial estimates of means and covariances, and the w-estimator equations are then iterated until convergence to obtain the S-estimates. An example shows that converging to a (local) minimum from the initial estimates from the elemental subsets is an effective way of determining the overall minimum. None of the estimates gained from the elemental samples is close to the final solution.

Biometry

Statistical procedures for estimating the detection limit and determination limit of the Ames Salmonella mutagenicity assay.

Novel and flexible procedures for estimating the detection limit as well as the determination limit of the Ames mutagenicity assay were proposed to evaluate the genetoxicity of a water sample. The accumulated data under the test conditions of TA 100-S9 by our group were taken as examples and analyzed to estimate the detection limit and the determination limit. The detection limit was estimated at 1.7 as the MR value when duplicate plates were used in the negative control test. However, it decreased to 1.4 as the MR level when quadruple plates were used in the negative control test. Therefore it was found that the sensitivity of the Ames mutagenicity assay was improved very easily by increasing the number of plates for the negative control test from two to four. The application of the conventional twofold rule to the data obtained with the strain TA100 was considered too conservative. The determination limit was regarded at 2.2 as the MR value under the following conditions: (a) quadruple plates were used in the negative control test; (b) three dose-steps including negative control step were designed at regular intervals; and (c) duplicate plates were used for each dose-step. It was proved by comparing data of two students that the detection limit and the determination limit estimated in this study were considered acceptable to any well trained students.

Biometry