PubMed Health⌕ Search

PubMed · 16026612

Oblique decision trees for spatial pattern detection: optimal algorithm and application to malaria risk.

Abstract

BACKGROUND: In order to detect potential disease clusters where a putative source cannot be specified, classical procedures scan the geographical area with circular windows through a specified grid imposed to the map. However, the choice of the windows' shapes, sizes and centers is critical and different choices may not provide exactly the same results. The aim of our work was to use an Oblique Decision Tree model (ODT) which provides potential clusters without pre-specifying shapes, sizes or centers. For this purpose, we have developed an ODT-algorithm to find an oblique partition of the space defined by the geographic coordinates. METHODS: ODT is based on the classification and regression tree (CART). As CART finds out rectangular partitions of the covariate space, ODT provides oblique partitions maximizing the interclass variance of the independent variable. Since it is a NP-Hard problem in RN, classical ODT-algorithms use evolutionary procedures or heuristics. We have developed an optimal ODT-algorithm in R2, based on the directions defined by each couple of point locations. This partition provided potential clusters which can be tested with Monte-Carlo inference. We applied the ODT-model to a dataset in order to identify potential high risk clusters of malaria in a village in Western Africa during the dry season. The ODT results were compared with those of the Kulldorff' s SaTScan. RESULTS: The ODT procedure provided four classes of risk of infection. In the first high risk class 60%, 95% confidence interval (CI95%) [52.22-67.55], of the children was infected. Monte-Carlo inference showed that the spatial pattern issued from the ODT-model was significant (p < 0.0001). Satscan results yielded one significant cluster where the risk of disease was high with an infectious rate of 54.21%, CI95% [47.51-60.75]. Obviously, his center was located within the first high risk ODT class. Both procedures provided similar results identifying a high risk cluster in the western part of the village where a mosquito breeding point was located. CONCLUSION: ODT-models improve the classical scanning procedures by detecting potential disease clusters independently of any specification of the shapes, sizes or centers of the clusters.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Jean Gaudart, Belco Poudiougou, Stéphane Ranque, Ogobara Doumbo. 2005-07-18. Oblique decision trees for spatial pattern detection: optimal algorithm and application to malaria risk.. https://doi.org/10.1186/1471-2288-5-22

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Diffuse cutaneous leishmaniasis in an HIV-positive patient in western Africa.

A 36-year-old HIV1-positive woman presented with a 6-month history of a progressive papular and nodular eruption of the face and subsequent extensive spread to the rest of the skin. The diagnosis of diffuse cutaneous leishmaniasis was established by direct examination and skin biopsy. This atypical form had a dramatic improvement after a 21-day treatment with meglumine antimoniate. This clinical form may be confused with other endemic diseases in western Africa, especially leprosy.

Africa, Western↗

Admixture in Mexico City: implications for admixture mapping of type 2 diabetes genetic risk factors.

Admixture mapping is a recently developed method for identifying genetic risk factors involved in complex traits or diseases showing prevalence differences between major continental groups. Type 2 diabetes (T2D) is at least twice as prevalent in Native American populations as in populations of European ancestry, so admixture mapping is well suited to study the genetic basis of this complex disease. We have characterized the admixture proportions in a sample of 286 unrelated T2D patients and 275 controls from Mexico City and we discuss the implications of the results for admixture mapping studies. Admixture proportions were estimated using 69 autosomal ancestry-informative markers (AIMs). Maternal and paternal contributions were estimated from geographically informative mtDNA and Y-specific polymorphisms. The average proportions of Native American, European and, West African admixture were estimated as 65, 30, and 5%, respectively. The contributions of Native American ancestors to maternal and paternal lineages were estimated as 90 and 40%, respectively. In a logistic model with higher educational status as dependent variable, the odds ratio for higher educational status associated with an increase from 0 to 1 in European admixture proportions was 9.4 (95%, credible interval 3.8-22.6). This association of socioeconomic status with individual admixture proportion shows that genetic stratification in this population is paralleled, and possibly maintained, by socioeconomic stratification. The effective number of generations back to unadmixed ancestors was 6.7 (95% CI 5.7-8.0), from which we can estimate that genome-wide admixture mapping will require typing about 1,400 evenly distributed AIMs to localize genes underlying disease risk between populations of European and Native American ancestry. Sample sizes of about 2,000 cases will be required to detect any locus that contributes an ancestry risk ratio of at least 1.5.

Africa, Western↗