PubMed HealthSearch

SEARCH · PubMed Health

Results for “Bayesian inference”

Explore indexed PubMed citations for clinical trials, systematic reviews and public health research. Read source abstracts and follow each citation to its original PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2Linked to original sources

Plastome evolution and phylogenomic relationships in Ajuga (Lamiaceae, Ajugoideae).

BACKGROUND: Ajuga is currently known to include approximately 69 species, with a combined distribution extending throughout Eurasia, Africa, and Australia. Its popularity and significance are largely based on an extensive history of medicinal and horticultural use. It is divided into two sections based on morphological characters, and this sectional classification is also reflected in pronounced geographic patterns. Although previous studies have largely focused on Ajuga sect. Ajuga in East Asia, A. sect. Chamaepithys, which ranges from the Mediterranean to Central Asia, remains insufficiently sampled, thereby limiting a comprehensive understanding of infrageneric sectional relationships within the genus. Here, we generated complete plastid genomes for 12 species representing both sections of the genus and used these data to characterize plastome structure and infer evolutionary relationships. RESULTS: In this study, 21 Ajuga plastomes were analyzed, including 12 newly sequenced plastomes and 9 previously published plastomes representing 19 species. Comparative analyses showed that all plastomes exhibited a highly conserved quadripartite structure, with genome sizes ranging from 149,963 to 150,740 bp and GC contents varying from 38.2% to 38.3%. Each plastome contained 133 genes, including 88 protein-coding genes, 37 transfer RNA genes, and 8 ribosomal RNA genes. The boundaries between the inverted repeat (IR) and single-copy (SC) regions were also highly conserved across species. In addition, 796 simple sequence repeats (SSRs), 874 long repeat sequences (LRSs), and 12 highly variable regions (ccsA-ndhD, ndhF-rpl32, petA-psbJ, rpl32-trnL-UAG, rps2-rpoC2, trnH-GUG-psbA, trnK-UUU-rps16, trnP-UGG-psaJ, trnT-UGU-trnL-UAA, ycf15-trnL-CAA, ndhF, and ycf1) were identified among the 21 plastomes. Phylogenetic analyses based on four datasets and conducted using Maximum Likelihood and Bayesian Inference recovered two major clades corresponding to the traditionally recognized sectional classification, with one distributed from the Mediterranean to Central Asia and the other in East Asia. CONCLUSION: This study represents the most comprehensive plastome-based sampling of Ajuga to date, including representative species from the Mediterranean, Central Asia, and East Asia. Our results have significantly enhanced our understanding of its infrageneric relationships. The plastome resources generated in this study provide a valuable foundation for future research on species delimitation, phylogeny, and the evolutionary history of Ajuga.

Phylogeny

Insights Into the Structural Features, Codon Usage Patterns, and Phylogenetic Analysis in Neoniphon argenteus (Teleostei: Holocentriformes) Based on Complete Mitochondrial Genome.

Neoniphon argenteus, a widely distributed nocturnal coral reef fish in the family Holocentridae, plays an important role in maintaining coral reef ecosystem health, yet its phylogenetic position remains poorly resolved. To bridge this gap, we sequenced and analyzed the complete mitochondrial genome of a specimen from the South China Sea to characterize its structural features, codon usage patterns, and phylogenetic relationships. The 16,569 bp mitogenome (GenBank: PP190474.1) encodes 13 protein-coding genes (PCGs), 22 tRNAs, two rRNAs, and two non-coding regions, exhibiting a distinct A + T bias. All tRNAs fold into typical cloverleaf secondary structures except tRNA-Ser (AGN), which lacks the dihydrouridine (DHU) arm. The control region contains palindromic motifs (TACAT/ATGTA) capable of forming hairpin structures and five conserved sequence blocks, whereas the OL region harbors a conserved 5'-GCCGG-3' motif. RSCU analysis revealed 31 frequently used codons (RSCU > 1) with a pronounced preference for A/C-ending codons. The ΔRSCU method identified 10 candidate optimal codons (GCA, CAA, GAA, GGA, AUU, CUA, CCA, CGA, ACA, and GUC). Selection pressure analysis using EasyCodeML and site-specific models indicated that all PCGs are predominantly under purifying selection, with no significant evidence of pervasive positive selection. ND6 exhibited elevated pairwise Ka/Ks ratios (mean = 1.209 ± 0.047), consistent with reduced selective constraint rather than adaptive evolution. Phylogenetic analysis of 19 Holocentriformes species using maximum likelihood and Bayesian inference with partitioned models based on 13 PCGs and two rRNA genes (12S and 16S) assigned all taxa to two well-supported subfamilies (Holocentrinae and Myripristinae). Within Holocentrinae, Neoniphon species form a monophyletic clade nested within a paraphyletic Sargocentron, suggesting that the genus Sargocentron as currently defined is not monophyletic. This study provides useful baseline molecular data for further exploration of the evolutionary history of N. argenteus and other members of Holocentriformes.

Holocentridae

Comparison of the information in two lung function experiments.

The amount of ventilation relative to perfusion (the ventilation-perfusion ratio) received by the lung is a useful indicator of the efficiency of lung function. Two alternative techniques for recovering the ventilation-perfusion ratio are outlined. While both techniques rely on the use of inert gases, one is well established and the other is only in a developmental stage. This paper focuses on a comparison of the amount of statistical information provided by these two techniques about the ventilation-perfusion ratio. The criterion applied here for measuring amount of information has roots in communication theory and uses ideas inherent to Bayesian inference.

Bayes Theorem

Probability and the patient state space.

This paper describes work to develop a model-based system to support clinical decision-making. In previous articles, we have developed (from 695 measurement sets obtained from 148 patients) a physiologic state classification based on a set of 11 cardiovascular and metabolic measurements. There is an R or reference state, for stable ICU patients. Patients under (operative, traumatic, or compensated septic) stress, or with (septic or hepatic) metabolic, respiratory, or cardiac insufficiency are in the A, B, C, or D states, respectively. We wished to make the state easier to measure and eventually available continuously, automatically, and noninvasively, as well as reflecting a wider group of bodily systems. The 5 centers define a 4 dimensional affine subspace, designated the cardiovascular state space. Using eigenvector analysis, we have found four new derived physiologic variables CV1, CV2, CV3, and CV4 that span the state space. We have fit sets of linear regression equations that allow the patient's position in the state space, and therefore his state, to be determined from more easily obtainable sets of measurements. Further, we selected 1966 measurement sets from 512 patients at two hospitals. We used the data from 250 of these patients to define 13 prototypical types, namely survivors and deaths from various combinations of sepsis, cardiogenic decompensation, cirrhosis, and pneumonitis, following trauma or general surgery. For any future patient, the statistical theory of Bayesian inference allows one to infer back from the measurements observed to the probability of his being of any of these types and of surviving or dying. We used this method to predict the outcome of the other 262 patients, prospectively. Statistically, the predictions of survival or death were not significantly different from the actual. For individual patients, the method predicts a clinical course that closely follows the actual episodes in their history. These results confirm and explain the validity of the concept of the patient state and make the state easier to compute. The patient state and the probability plot together help to stage, select, and evaluate therapy. They do not replace the clinician's judgement, but rather are tools that help the clinician to exercise judgement.

Adult

Medical expert systems based on causal probabilistic networks.

Causal probabilistic networks (CPNs) offer new methods by which you can build medical expert systems that can handle all types of medical reasoning within a uniform conceptual framework. Based on the experience from a commercially available system and a couple of large prototype systems, it appears that CPNs are now an attractive alternative to other methods. A CPN is an intensional model of a domain, and it is therefore conceptually much closer to qualitative reasoning systems and to simulation systems than to rule-based or logic-based systems. Recent progress in Bayesian inference in networks has yielded computationally efficient methods. The inference method used follows the fundamental axioms of probability theory, and gives a sound framework for causal and diagnostic (deductive and abductive) reasoning under uncertainty. Experience with the prototypes indicates that it may be possible to use decision theory as a rational approach to test planning and therapy planning. The way in which knowledge is acquired and represented in CPNs makes it easy to express 'deep knowledge' for example in the form of physiological models, and the facilities for learning make it possible to make a smooth transition from expert opinion to statistics based on empirical data.

Artificial Intelligence

Mitochondrial genome characteristics and phylogenetic analysis of Ramaria longispora.

This study, for the first time, assembled and annotated the complete mitochondrial genome of R. longispora using high-throughput sequencing technology. The genome is a circular molecule with a total length of 157,712 bp and a GC content of 31.55%. It encodes 71 genes, including 15 core protein-coding genes (PCGs), 25 transfer RNA (tRNA) genes, 2 ribosomal RNA (rRNA) genes, 5 free-stranding open reading frames (ORFs), and 24 intronic ORFs. Among these, most free-stranding ORFs have unknown functions but include a DNA polymerase gene, while the intronic ORFs primarily encode LAGLIDADG and GIY-YIG endonucleases. The mitochondrial genome contains 39 introns. Phylogenetic analyses based on 15 core PCGs using Bayesian inference (BI) and maximum likelihood (ML) methods revealed that this R. longispora is most closely related to Ramaria flavescens and Ramaria ichnusensis. This study provides foundational data for mitochondrial genome research in the Ramaria genus and offers important references for taxonomic and evolutionary studies of this group.

Mitochondrial genome

GAMMA: gap-aware motif mining under incomplete labeling with applications to MHC motifs.

MOTIVATION: Sequence motif identification is crucial for understanding molecular recognition, particularly in immune responses involving peptide binding to major histocompatibility complex (MHC) Class I molecules for antigen presentation to T cells. Traditionally, MHC Class I binding motifs are assumed to be contiguous and span nine amino acids. However, structural evidence suggests that binding may involve nonadjacent residues, challenging the assumptions of existing methods. RESULTS: In this study, we propose Gap-Aware Motif Mining Algorithm (GAMMA), a probabilistic framework designed to identify noncontiguous motifs under conditions of incomplete labeling. GAMMA employs Bayesian inference with Markov chain Monte Carlo sampling to jointly estimate motif parameters, binding locations, and the relative spacing between binding positions. Through extensive simulations and real-world applications to MHC Class I peptide datasets, GAMMA outperforms existing motif discovery tools such as GLAM2 in accurately localizing binding residues and identifying the underlying motifs. Notably, our results suggest that the true number of binding residues may be eight, fewer than the commonly assumed nine. In addition, for longer peptides, the model captures increased flexibility in the central region, consistent with structural observations that peptides may bulge in the middle. AVAILABILITY AND IMPLEMENTATION: The raw data and the source codes are available on GitHub (https://github.com/RanLIUaca/GAMMAmotif).

Amino Acid Motifs

Detecting Interspecific Positive Selection Using Convolutional Neural Networks.

Traditional statistical methods using maximum likelihood and Bayesian inference can detect positive selection from an interspecific phylogeny and a codon sequence alignment based on model assumptions, but they are prone to false positives due to alignment errors and can lack power. These problems are particularly pronounced when faced with high levels of indels and divergence. To address these issues, we trained and tested convolutional neural network models on simulated data and achieved higher accuracy in detecting selection across a specific range of phylogenetic scenarios and evolutionary modes. This advantage is particularly evident when performing inference on noisy data prone to misalignments. Our method shows some ability to account for these errors, where most statistical frameworks fail to do so in a tractable manner. We explore the generalizability of our convolutional neural network models to unseen evolutionary scenarios and identify future avenues to achieve broader utility. Once trained, our convolutional neural network model is faster at test time, making it a scalable alternative to traditional statistical methods for large-scale, multigene analyses. In addition to binary classification (inference of the presence or absence of positive selection during the evolution of the sequences), we use saliency maps to understand what the model learns and observe how this could be leveraged for sitewise inference of positive selection.

Neural Networks, Computer

Mapping Sub-National Respiratory Virus Circulation in Cambodia Using Metatranscriptomic Sequencing: A Multi-Center Hospital-Based Surveillance Study.

BACKGROUND: Genomic surveillance can guide early detection of and response to emerging epidemics. Metatranscriptomic sequencing was used to investigate sub-national respiratory virus circulation in Cambodia from 2020 to 2023. METHODS: Nasopharyngeal swabs were collected from individuals aged 2 months to 65 years with influenza-like illness in four Cambodian hospitals. Metatranscriptomic data were generated by short-read RNA sequencing. Bernoulli space-time scan statistics were used to identify temporal virus clusters. Bayesian inference of phylogenetic trees was used to compute divergence times for temporally clustered, highly represented viruses (influenza A/H3N2 and B, Betacoronavirus 1, respiratory syncytial virus [RSV] A and B), and publicly available global influenza virus genomes. RESULTS: Of 1093 individuals, 499 (45.7%) had detectable respiratory viruses belonging to 68 distinct species. Moderate (N > 20) discrete time-clusters were noted of RSV-A (37 cases), Betacoronavirus 1 (21 cases), RSV-B (22 cases), and A/H3N2 (30 cases). The posterior median of time to most recent common ancestor ranged from 0.71 years (95% HPD 0.38-1.10) for Betacoronavirus 1 and 1.31 years (95% HPD 0.60-3.20) for A/H3N2, to 2.75 years (1.82-4.26) for RSV-A and 4.79 years (2.39-7.74) for RSV-B. A/H3N2 and influenza B virus genomes mapped to clades 3C.2a1b.2a.2a and Victoria 1A.3a.2, respectively, and inter-mixed with concurrent global strains. CONCLUSIONS: Multiple respiratory viruses circulated at a sub-national level in Cambodia from 2020 to 2023 despite pandemic disruptions. Influenza virus population diversity decreased during the height of lockdown but recovered in mid-2022. Re-emerging influenza strains were distinct from historically circulating strains and clustered with contemporaneous global variants, suggesting multiple external introductions.

Humans

On avoiding statistical bias in linkage-based counselling.

Using the Succession Rule of Laplace (1795) and related reasoning, this paper shows how to give unbiased counselling to patients when predictions are to be based on small samples. The recombination fraction can be regarded as a probability parameter, theta, which itself has a probability distribution between the limits of 0 and 1/2. The probability of a recombinant, P(Rec), is not numerically equal to the maximum likelihood estimate of theta, nor is it numerically equal to the maximum posterior probability estimate in Bayesian inference. Rather it is equal to the infinite sum of all possible theta values, each weighted according to its probability density p(theta) which denotes the relative probability that that theta value is the true one. The various published proposals for obtaining an unbiased estimate of theta are shown to be equivalent one to another, except for the simplifying approximations used.

Bias

Contrasting Patterns of Connectivity Between Populations of Euphotic and Mesophotic Hydroids in Reunion Island Support the Deep Reef Refuge Hypothesis.

In the context of coral reef decline, mesophotic coral ecosystems (MCEs, 30-150 m) offer hope for the recovery of degraded euphotic reefs. The Deep Reef Refuge Hypothesis (DRRH) postulates the potential of mesophotic reefs to reseed euphotic reefs. This hypothesis needs to be further tested by estimating connectivity along the depth gradient. Mesophotic data are lacking worldwide, particularly in the southwestern Indian Ocean (SWIO). Here, using a total of 2218 samples collected at depths ranging from 10 to 103 m, we estimated the connectivity of 7 hydroid species sampled at euphotic, upper, and lower mesophotic depths around Reunion Island using a multi-species comparative framework. Population genetic analyses using 8-17 microsatellite markers per species (80 markers in total) as well as Bayesian inference were performed to estimate population structure and contemporary migration rates to highlight connectivity patterns and directionality of gene flow between depths. The results revealed three main genetic patterns depending on the species: a horizontal stepping stone pattern between areas around the island, a vertical stepping stone pattern between adjacent depths, and a quasi-panmictic pattern. Each species showed some specificity within these patterns, but overall, at least 4 of the 7 species support the assumption of vertical connectivity from the Deep Reef Refuge Hypothesis, highlighting the importance of studying multiple species. The existence of vertical connectivity between euphotic and mesophotic depths in the southwestern Indian Ocean confirms the importance of mesophotic coral ecosystems for conservation efforts and our global understanding of coral reef ecosystem dynamics.

Animals

Experiencing and perceiving visual surfaces.

A theoretical framework is proposed to understand binocular visual surface perception based on the idea of a mobile observer sampling images from random vantage points in space. Application of the generic sampling principle indicates that the visual system acts as if it were viewing surface layouts from generic not accidental vantage points. Through the observer's experience of optical sampling, which can be characterized geometrically, the visual system makes associative connections between images and surfaces, passively internalizing the conditional probabilities of image sampling from surfaces. This in turn enables the visual system to determine which surface a given image most strongly indicates. Thus, visual surface perception can be considered as inverse ecological optics based on learning through ecological optics. As such, it is formally equivalent to a degenerate form of Bayesian inference where prior probabilities are neglected.

Depth Perception

Effect of atypical antibiotic resistance on microorganism identification by pattern recognition.

We classified microorganisms from the clinical laboratory by using information provided by the Gram stain and antibiotic sensitivity profiles obtained with the Bauer-Kirby technique. Approximately 4,000 microorganisms, routinely identified and tested for antibiotic sensitivities in a large hospital microbiology laboratory, were used as a data set for several pattern recognition classification methods: K--nearest-neighbor analysis, statistical isolinear multicomponent analysis, Bayesian inference, and linear discriminant analysis. K--nearest-neighbor analysis yielded the highest prospective classification accuracy for gram-negative organisms, 90%. When those organisms displaying an atypical antibiotic resistance pattern were excluded from the data, the gram-negative classification accuracy improved to 95%. These results are inferior to currently accepted biochemical identification methods. Microorganisms with atypical antibiotic resistance patterns are likely to be misidentified and are common enough (17% of our isolates) to limit the feasibility of routine identification of microorganisms from their antibiotic sensitivities.

Anti-Bacterial Agents

The risk of falling in the elderly: a subjective approach.

Accidental injury resulting from falling is the leading cause of death in the elderly population. In fact, 11% of the nation's annual injury deaths occur in this age group. A significant number of accidental deaths occur in resident institutions. A short-term risk index was constructed and validated for the purpose of predicting the risk of a fall in an elderly institutionalized population. The risk index was developed from expert insights, using Bayesian inference. A group of experts identified major risk factors that influence the chance of a fall and provided estimates of the impacts of the identified risk factors. Identification of risk factors, elicitation of experts' insights, construction of the subjective risk index, and validation of the index are discussed.

Accidental Falls

Evolutionary dynamics of the chloroplast genome in Abutilon (Malvoideae, Malvaceae).

The genus Abutilon Mill. (Malvaceae) comprises approximately 178 species distributed across tropical and subtropical regions, many of which hold significant ornamental, economic, and medicinal value; yet its taxonomic classification remains challenging. In this study, six species were sequenced from herbarium specimens, and the chloroplast (cp.) genomes of ten additional species were assembled de novo from publicly available raw data. Three previously reported cp. genomes were also incorporated to characterise cp. genome structure, identify polymorphic loci, and perform phylogenetic analyses. The cp. genomes ranged from 159,458 to 160,454 bp and exhibited the typical quadripartite structure, with each genome containing 112 unique genes (78 protein-coding, 30 tRNA, and 4 rRNA) that showed conserved content and organisation. These genomes exhibited high similarity in GC content, inverted repeat boundaries, relative synonymous codon usage, amino acid frequencies, and substitution patterns. However, notable variation was observed in the total number of simple sequence repeats, ranging from 70 to 97 per genome. Selection analyses indicated predominant purifying selection, with evidence of episodic positive selection detected in rpoC2, rbcL, and ycf1. Two codons in rbcL were clade-specific and provided phylogenetic signal distinguishing Australian and Old World pantropical species. Nucleotide diversity analysis identified six highly polymorphic intergenic spacers (trnH-psbA, rps19-rpl2, psbT-pbf1, psaC-ndhD, trnR-atpA, and ndhJ-ndhK) that may be suitable for taxonomic studies. The phylogeny from maximum likelihood (ML) and Bayesian inference (BI) resolved two major clades: one comprising an exclusively Australian lineage occurring predominantly in arid and semi-arid environments, and the other a pantropical lineage spanning multiple continents. Abutilon grandifolium was recovered as sister to the remaining sampled Abutilon taxa in both ML and BI analyses, although no biogeographic origin inference can be drawn from this placement pending broader taxon sampling and integration of nuclear genomic data. These findings provide insights into the evolutionary dynamics of the cp. genome in Abutilon and offer a foundational genomic framework for refining Abutilon taxonomy.

Genome, Chloroplast

The formation of maintenance of delusions: a Bayesian analysis.

This paper argues that recent research on normal-belief formation is relevant to our understanding of the establishment and maintenance of delusions. Bayesian theory provides a normative model of the way in which evidence relevant to normal beliefs may be evaluated: this makes it possible to classify delusional beliefs in terms of deviations from optimal Bayesian inference. Some hypothetical forms of deviation appear to correspond closely to cognitive processes observed in some groups of deluded patients. Theories of the precise nature of the abnormal judgemental processes also have implications for psychological approaches to treatment of deluded patients. The role of hallucinations in the formation and/or maintenance of delusions and the extent to which the distortions of cognitive processes associated with delusions are content-specific or mood-specific are also considered.

Cognition Disorders

Pan genome clustering identifies a novel mosaic prophage specific to Salmonella Enteritidis lineage associated with the invasive disease in India.

Salmonella enterica serovar Enteritidis is a leading cause of invasive non-typhoidal Salmonella (iNTS) disease globally, particularly in sub-Saharan Africa. In contrast, the epidemiology and population structure of invasive S. Enteritidis in South Asia remain poorly characterized. This study investigates the clinical presentation, phylogenetic relationships and genomic characteristics of S. Enteritidis bloodstream infections (BSIs) in India. Clinical data were collected from 101 patients with S. Enteritidis BSI between 2012 and 2022. Whole-genome sequencing was performed on representative bloodstream isolates together with isolates from non-blood clinical specimens and poultry sources. Comparative genomic analyses included phylogenetic reconstruction, invasiveness index prediction, and prophage characterization. Infants and immunosuppressed individuals were disproportionately affected by iNTS disease. Phylogenetic analysis identified four major lineages of S. Enteritidis. Most BSI isolates clustered in a previously unrecognized lineage, designated the Global Intermediate Clade, which occupied a phylogenetic position between the Global outlier and Global epidemic clades. Bayesian inference dated its most recent common ancestor to around 1789 AD (95% HPD: 1692-1941), with global circulation confirmed by European and Asian isolates. The Global Intermediate clade exhibited the second-highest invasiveness index (median 0.221, SD 0.013) after the West African clade; however, this index reflects genomic signatures associated with invasiveness and should not be interpreted as a direct measure of virulence. Poultry isolates clustered separately from the dominant bloodstream-associated lineage. Pan-genome analysis identified a lineage-specific mosaic prophage composed of modules homologous to prophages found in diverse Enterobacterales. This study provides the first detailed genomic insight into invasive S. Enteritidis in India and identifies a previously unrecognized Global Intermediate Clade associated with bloodstream infection. The distinct phylogenetic placement and genomic features of this lineage, including a lineage-specific mosaic prophage, warrant further investigation and support the need for expanded One Health genomic surveillance.

Humans

Genomic background of gestation length and calving-related traits in Holstein cattle.

The reproductive success of cows directly influences the profitability of dairy farms. Reproductive traits, particularly calving-related traits, generally have low heritability but sufficient additive genetic variance to enable genetic progress through genomic selection. Thus, the primary objectives of this study were to estimate genetic parameters and perform single-step genome-wide association studies (ssGWAS) for calf size, calving ease, gestation length, and stillbirth in Holstein cattle. Variance components were estimated based on animal models and Bayesian inference using a data set containing 226,717 animals with phenotypic records, 15,761 animals genotyped with 45,101 SNP markers, and 461,819 animals in the pedigree. SNP effects were estimated using the single-step GBLUP method. For direct and maternal genetic effects, heritability estimates (posterior standard deviation) ranged from 0.001 (0.002) for gestation length in heifers to 0.16 (0.001) for gestation length in cows. Genetic correlations ranged from -0.57 (0.01) between calving ease and stillbirth in heifers to 0.74 (0.01) between gestation length evaluated in heifers and cows. The ssGWAS results supported a highly polygenic architecture for calving-related traits, with most genomic signals not reaching genome-wide significance. A genome-wide significant association was detected for calving ease in cows on BTA23, highlighting FARS2 as a positional candidate gene. The strongest GWAS signals for each trait harbored additional biologically important candidate genes, including NPPA, NPPB, BCHE, EPHA4, DLD, and GTF2I. Given the generally low heritability estimates and the predominantly polygenic architecture observed for these traits, genomic selection may contribute to the genetic improvement of calving-related traits in Holstein cattle, with potential benefits for cow welfare, calf survival, and overall dairy production efficiency.

dairy cattle