PubMed Health⌕ Search

PubMed · 16408921

Improving the linearity of infrared diffuse reflection spectroscopy data for quantitative analysis: an application in quantifying organophosphorus contamination in soil.

Abstract

Diffuse reflection data are presented for ethyl methylphosphonate in a fine Utah dirt sample as a model system for organophosphate-contaminated soil. The data revealed a chemometric artifact when the spectra were represented in Kubelka-Munk units that manifests as a linear dependence of spectral peak height on variations in the observed baseline position (i.e., the position of the observed transmission intensity where no absorption features occur in the sample spectrum). We believe that this artifact is the result of the mathematical process by which the raw data are converted into Kubelka-Munk units, and we developed a numerical strategy for compensating for the observed effect and restoring chemometric precision to the diffuse reflection data for quantitative analysis while retaining the benefits of linear calibration afforded by the Kubelka-Munk approach. We validated our Kubelka-Munk correction strategy by repeating the experiment using a simpler system--pure caffeine in potassium bromide. The numerical preprocessing includes conventional multiplicative scatter correction coupled with a baseline offset correction that facilitates the use of quantitative diffuse reflection data in the Kubelka-Munk formalism for the quantitation of contaminants in a complex soil matrix, but is also applicable to more fundamental diffuse reflection quantitative analysis experiments.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Alan C Samuels, Changjiang Zhu, Barry R Williams, Avishai Ben-David, Ronald W Miles, Melissa Hulet. 2006-01-15. Improving the linearity of infrared diffuse reflection spectroscopy data for quantitative analysis: an application in quantifying organophosphorus contamination in soil.. https://doi.org/10.1021/ac0509859

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

A robust transfer learning approach for high-dimensional linear regression to support integration of multi-source gene expression data.

Transfer learning aims to integrate useful information from multi-source datasets to improve the learning performance of target data. This can be effectively applied in genomics when we learn the gene associations in a target tissue, and data from other tissues can be integrated. However, heavy-tail distribution and outliers are common in genomics data, which poses challenges to the effectiveness of current transfer learning approaches. In this paper, we study the transfer learning problem under high-dimensional linear models with t-distributed error (Trans-PtLR), which aims to improve the estimation and prediction of target data by borrowing information from useful source data and offering robustness to accommodate complex data with heavy tails and outliers. In the oracle case with known transferable source datasets, a transfer learning algorithm based on penalized maximum likelihood and expectation-maximization algorithm is established. To avoid including non-informative sources, we propose to select the transferable sources based on cross-validation. Extensive simulation experiments as well as an application demonstrate that Trans-PtLR demonstrates robustness and better performance of estimation and prediction when heavy-tail and outliers exist compared to transfer learning for linear regression model with normal error distribution. Data integration, Variable selection, T distribution, Expectation maximization algorithm, Genotype-Tissue Expression, Cross validation.

Linear Models↗

Quantifying the uncertainty of on-line sensors at WWTPs during field operation.

It remains an ongoing task to quantify the uncertainty of continuous measuring systems at WWTPs during field operation. The commonly used methods are based on lab experiments under standardized conditions and are only suitable for characterizing the measuring device itself. For measuring devices under field conditions, a knowledge of the response time, trueness and precision is equally important. A method is proposed which can be used to characterize newly installed on-line sensors or to evaluate monitoring data which may contain systematic errors. The concept is based on comparative measurements between the sensor and a reference. A linear regression is used to differentiate between trueness and precision. Various statistical tests are conducted to validate the preconditions of linear regression. The information about the trueness and precision of the measuring system under field conditions helps to adapt control strategies more effectively to the relevant processes and permits sophisticated control concepts. Moreover, the concept can help to define guidelines for evaluating the uncertainties of effluent quality monitoring to overcome the concerns about on-line sensors, improve the trust in these systems and to allow the use of continuously measuring systems for legislative purposes. The approach is discussed in detail in this paper and all statistical tests and formulas are listed in the Appendix.

Linear Models↗

Linear relationship between activation energies and reaction energies for coverage-dependent dissociation reactions on rhodium surfaces.

Evidence of a relationship between activation energies and enthalpy changes of various dissociation reactions on transition metals has been reported recently. A reconsideration of density functional theory results for dissociation energies of oxygen and NO on different rhodium surfaces (low-index and stepped) and their dependencies on oxygen precoverage reveal that also here a linear Brønsted-Evans-Polanyi (BEP) relationship exists. The establishment of such a general concept would be of tremendous importance for the development of detailed, elementary-step reaction mechanisms, because the activation energies of reaction steps as well as their coverage dependencies could be estimated based on the adsorption energies calculated by means of DFT.

Linear Models↗