bioRxiv ScienceSearch

bioRxiv · 10.1101/070144

Analysis of infection biomarkers within a Bayesian framework reveals their role in pneumococcal pneumonia diagnosis in HIV patients

Abstract

BackgroundHIV patients are more likely to contract bacterial pneumonia and more likely to die from the infection. Unfortunately, there are few tests to quickly diagnosis the etiology of these dangerous infections. Several biomarkers may be useful for diagnosing the most common pneumonia-causing organism, S. pneumoniae, but studies utilizing the standard statistical approach provide little concrete guidance for the HIV-infected population.\n\nMethodology and FindingsUsing a Bayesian approach, I analyze data from a cohort of 280 HIV patients with x-ray confirmed community acquired pneumonia. First, I use a variety of techniques to establish predictor significance and to identify their optimal cutoffs. Next, in lieu of cutoffs, I find the continuous and combined likelihood ratios for every value of each biomarker, and I compute the associated posttest probabilities. As expected, I find the three biomarkers with good clinical yield and a statistically significant association with S. pneumoniae are C-reactive protein (CRP), procalcitonin (PCT), and lytA gene PCR (lytA). Based on Bayesian clinical yield, optimal cutoffs are largely equivocal. The optimal dichotomous cutoff for CRP is essentially any value between 10 mg/dL and 30 mg/dL ({bigtriangleup}pPosttest {approx} 0.49). The optimal cutoff for PCT is any value between 2 ng/mL and 40 ng/mL ({bigtriangleup}pposttest {approx} 0.35). The optimal cutoff for lytA is any value less than 6 log10 copies/mL ({bigtriangleup}pposttest {approx} 0.45). Further, I find that continuous likelihood ratios provide much more accurate posttest probabilities than dichotomous cutoffs. For example, starting with the empirical pretest probability, a lytA approaching 0 copies/mL lowers the probability of S. pneumoniae infection to less than 15%, while a result of 10 copies/mL raises the probability to greater than 65%. However, a lytA value just above or below the suggested cutoff of 8000 copies/mL or my new optimal cutoff of 30,000 copies/mL leaves the posttest probability of infection essentially unchanged from the pretest probability.\n\nConclusionCRP, PCT, and lytA all provide significant value in diagnosing the etiology of pneumonia in HIV patients. The optimal dichotomous cutoffs for lytA, CRP, and PCT need to be adjusted for pneumococcal diagnosis in this population. However, continuous and combined likelihood ratios avoid discarding valuable quantitative information, and a combined likelihood ratio can be easily computed without the need for prior logistic regression. Importantly, there is significant overlap between these biomarkers such that only one of the three biomarkers at a time should be used to update clinical probabilities. Thus, it is ill-advised to combine the likelihood ratios of different biomarkers to produce a posttest probability. Finally, I provide a simple web application to quantitatively calculate the posttest probability of S. pneumoniae infection in HIV patients with x-ray confirmed pneumonia: http://meyerapps.org/pneumococcal_etiology_hiv.

Explore related subjects

Keep this discovery

BibTeXRIS

Austin G Meyer. 2016-08-18. Analysis of infection biomarkers within a Bayesian framework reveals their role in pneumococcal pneumonia diagnosis in HIV patients. https://doi.org/10.1101/070144

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

A risk stratification approach for improved interpretation of diagnostic accuracy statistics

Diagnostic accuracy statistics, including predictive values, risk-differences, Youdens index and Area Under the Curve (AUC), assess the promise of novel biomarkers proposed as diagnostic tests. We reinterpret these statistics in light of risk-stratification (how well a biomarker separates those at higher risk from those at lower risk) to better understand their implications for public-health programs. We introduce an intuitively simple statistic, Mean Risk Stratification (MRS): the average change in risk (pre-test vs. post-test) revealed for tested individuals. High MRS implies better risk separation achieved by testing. MRS demonstrates that conventional predictive values can mislead because they do not account for disease prevalence or test-positivity rates. Little risk-stratification is possible for rare diseases, demonstrating a \"high-bar\" to justify population-based screening. Importantly, we demonstrate that the risk-difference, Youdens index, and AUC measure only multiplicative relative gains in risk-stratification: AUC=0.6 achieves only 20% of maximum risk-stratification (AUC=0.9 achieves 80%). However, large relative gains in risk-stratification might not imply large absolute gains if disease is rare or if the test is rarely positive. We illustrate MRS by our experience comparing the performance of cervical cancer screening tests in China vs. the USA. The test with the worst AUC=0.72 in China (visual inspection with ascetic acid) provides twice the risk-stratification of the test with best AUC=0.83 in the USA (human papillomavirus and Pap cotesting) because China has three times more cervical precancer/cancer. MRS could be routinely calculated to better understand the clinical/public-health implications of standard diagnostic accuracy statistics.

Epidemiology

Collider Scope: How selection bias can induce spurious associations

Large-scale cross-sectional and cohort studies have transformed our understanding of the genetic and environmental determinants of health outcomes. However, the representativeness of these samples may be limited - either through selection into studies, or by attrition from studies over time. Here we explore the potential impact of this selection bias on results obtained from these studies, from the perspective that this amounts to conditioning on a collider (i.e., a form of collider bias). While it is acknowledged that selection bias will have a strong effect on representativeness and prevalence estimates, it is often assumed that it should not have a strong impact on estimates of associations. We argue that because selection can induce collider bias (which occurs when two variables independently influence a third variable, and that third variable is conditioned upon), selection can lead to substantially biased estimates of associations. In particular, selection related to phenotypes can bias associations with genetic variants associated with those phenotypes. In simulations, we show that even modest influences on selection into, or attrition from, a study can generate biased and potentially misleading estimates of both phenotypic and genotypic associations. Our results highlight the value of knowing which population your study sample is representative of. If the factors influencing selection and attrition are known, they can be adjusted for. For example, having DNA available on most participants in a birth cohort study offers the possibility of investigating the extent to which polygenic scores predict subsequent participation, which in turn would enable sensitivity analyses of the extent to which bias might distort estimates.\n\nKey MessagesSelection bias (including selective attrition) may limit the representativeness of large-scale cross-sectional and cohort studies.\n\nThis selection bias may induce collider bias (which occurs when two variables independently influence a third variable, and that variable is conditioned upon).\n\nThis may lead to substantially biased estimates of associations, including of genetic associations, even when selection / attrition is relatively modest.

Epidemiology

A comparative analysis of Chikungunya and Zika transmission

The recent global dissemination of Chikungunya and Zika has fostered public health concern worldwide. To better understand the drivers of transmission of these two arboviral diseases, we propose a joint analysis of Chikungunya and Zika epidemics in the same territories, taking into account the common epidemiological features of the epidemics: transmitted by the same vector, in the same environments, and observed by the same surveillance systems. We analyse eighteen outbreaks in French Polynesia and the French West Indies using a hierarchical time-dependent SIR model accounting for the effect of virus, location and weather on transmission, and based on a disease specific serial interval. We show that Chikungunya and Zika have similar transmission potential in the same territories (transmissibility ratio between Zika and Chikungunya of 1.04 [95% credible interval: 0.97; 1.13]), but that detection and reporting rates were different (around 19% for Zika and 40% for Chikungunya). Temperature variations between 22{degrees}C and 29{degrees}C did not alter transmission, but increased precipitation showed a dual effect, first reducing transmission after a two-week delay, then increasing it around five weeks later. The present study provides valuable information for risk assessment and introduces a modelling framework for the comparative analysis of arboviral infections that can be extended to other viruses and territories.

Epidemiology