PubMed Health⌕ Search

PubMed · 11464753

Correlations between microbial parameters from water samples: expectations and reality.

Abstract

Data which are collected in order to estimate the correlation between parameters must be analysed with caution. Classical statistics of correlation are often inappropriate. The "r" statistic is very easily distorted by non-Normal data. Non-parametric statistics can be helpful. The interpretation and usefulness of the estimates of correlation will depend on the study plan. If water samples come from disparate sources (e.g. upstream or downstream from sewage outlets) then parameters A and B may occur in their highest and lowest numbers according to how close the samples were to contamination sources thus correlating closely. However, if all samples come from sources with similar pollution levels then plots of A and B will show considerable scatter and apparently little correlation. So what is the relationship between A and B? An example of "perfect" correlation, as demonstrated by replicate counts of a single parameter from split samples, gave an r value of only 0.63 (p = 0.62) due to random variation in numbers of organisms between the two halves of the sample. Thus large amounts of data are needed for studying true correlation because relationships between parameters are embedded in the natural variation. This also illustrated that Standards for a single parameter can be "passed" or "failed" by two halves of the same sample. Study design is clearly of fundamental importance. Consideration must be given to the appropriate way of asking questions about correlation between different parameters.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

H E Tillett, J Sellwood, N F Lightfoot, P Boyd, S Eaton. 2001. Correlations between microbial parameters from water samples: expectations and reality.. https://pubmed.ncbi.nlm.nih.gov/11464753/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Probability estimation when some observations are grouped.

This paper considers the use of additional questions for decreasing survey non-response rates and an approach for estimating a probability based on the results obtained. In a survey, the respondents are asked to answer an original question and follow-up questions, where the answers for the follow-up questions are grouped answers for the original question. For example, respondents are asked to provide an exact number of incidents, but in cases of 'Do not know' or 'Refuse' responses, they are subsequently asked to pick an answer from a less specific categorical scale. The new estimator obtains smaller variance asymptotically and does not depend on a distribution family. This method is applied to income questions in a survey regarding injury prevention and behaviours. Another application is survey data on intimate partner violence, where some amendments were applied for incorporating post-stratification weights and for using non-random grouping. For additional illustration, an example of parameter estimation on artificially generated data is presented.

Data Collection↗