bioRxiv ScienceSearch

bioRxiv · 10.1101/186387

Clustering of adult-onset diabetes into novel subgroups guides therapy and improves prediction of outcome

Abstract

BackgroundDiabetes is presently classified into two main forms, type 1 (T1D) and type 2 diabetes (T2D), but especially T2D is highly heterogeneous. A refined classification could provide a powerful tool individualize treatment regimes and identify individuals with increased risk of complications already at diagnosis.\n\nMethodsWe applied data-driven cluster analysis (k-means and hierarchical clustering) in newly diagnosed diabetic patients (N=8,980) from the Swedish ANDIS (All New Diabetics in Scania) cohort, using five variables (GAD-antibodies, BMI, HbA1c, HOMA2-B and HOMA2-IR), and related to prospective data on development of complications and prescription of medication from patient records. Replication was performed in three independent cohorts: the Scania Diabetes Registry (SDR, N=1466), ANDIU (All New Diabetics in Uppsala, N=844) and DIREVA (Diabetes Registry Vaasa, N=3485). Cox regression and logistic regression was used to compare time to medication, time to reaching the treatment goal and risk of diabetic complications and genetic associations.\n\nFindingsWe identified 5 replicable clusters of diabetes patients, with significantly different patient characteristics and risk of diabetic complications. Particularly, individuals in the most insulin-resistant cluster 3 had significantly higher risk of diabetic kidney disease, but had been prescribed similar diabetes treatment compared to the less susceptible individuals in clusters 4 and 5. The insulin deficient cluster 2 had the highest risk of retinopathy. In support of the clustering, genetic associations to the clusters differed from those seen in traditional T2D.\n\nInterpretationWe could stratify patients into five subgroups predicting disease progression and development of diabetic complications more precisely than the current classification. This new substratificationn may help to tailor and target early treatment to patients who would benefit most, thereby representing a first step towards precision medicine in diabetes.\n\nFundingThe funders of the study had no role in study design, data collection, analysis, interpretation or writing of the report.\n\nResearch in contextEvidence before this study\n\nThe current diabetes classification into T1D and T2D relies primarily on presence (T1D) or absence (T2D) of autoantibodies against pancreatic islet beta cell autoantigens and age at diagnosis (earlier for T1D). With this approach 75-85% of patients are classified as T2D. A third subgroup, Latent Autoimmune Diabetes in Adults (LADA,<10%), is defined by presence of autoantibodies against glutamate decarboxylase (GADA) with onset in adult age. In addition, several rare monogenic forms of diabetes have been described, including Maturity Onset Diabetes of the Young (MODY) and neonatal diabetes. This information is provided by national guidelines (ADA,WHO, IDF, Diabetes UK etc) but has not been much updated during the past 20 years and very few attempts have been made to explore heterogeneity of T2D. A topological analysis of potential T2D subgroups using electronic health records was published in 2015 but this information has not been implemented in the clinic.\n\nAdded value of this study\n\nHere we applied a data-driven cluster analysis of 5 simple variables measured at diagnosis in 4 independent cohorts of newly-diagnosed diabetic patients (N=14755) and identified 5 replicable clusters of diabetes patients, with significantly different patient characteristics and risk of diabetic complications. Particularly, individuals in the most insulin-resistant cluster 3 had significantly higher risk of diabetic kidney disease.\n\nImplications of the available evidence\n\nThis new sub-stratification may help to tailor and target early treatment to patients who would benefit most, thereby representing a first step towards precision medicine in diabetes

Explore related subjects

Keep this discovery

BibTeXRIS

Ahlqvist, E., Storm, P., Karajamaki, A., Martinell, M., Dorkhan, M., Carlsson, A., Vikman, P., Prasad, R. B., Mansour Aly, D., Almgren, P., Wessman, Y., Shaat, N., Spegel, P., Mulder, H., Lindholm, E., Melander, O., Hansson, O., Malmqvist, U., Lernmark, A., Lahti, K., Forsen, T., Tuomi, T., Rosengren, A. H., Groop, L.. 2017-09-08. Clustering of adult-onset diabetes into novel subgroups guides therapy and improves prediction of outcome. https://doi.org/10.1101/186387

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Translating surveillance data into incidence estimates

Monitoring a population for a disease requires the hosts to be sampled and tested for the pathogen. This results in sampling series from which to estimate the disease incidence, i.e. the proportion of hosts infected. Existing estimation methods assume that disease incidence is not changing between monitoring rounds, resulting in underestimation of the disease incidence. In this paper we develop an incidence estimation model accounting for epidemic growth with monitoring rounds sampling varying incidence. We also show how to accommodate the asymptomatic period characteristic to most diseases. For practical use, we produce an approximation of the model, which is subsequently shown accurate for relevant epidemic and sampling parameters. Both the approximation and the full model are applied to stochastic spatial simulations of epidemics. The results prove their consistency for a very wide range of situations.

epidemiology

The Swiss Primary Ciliary Dyskinesia registry: objectives, methods and first results

Primary Ciliary Dyskinesia (PCD) is a rare hereditary, multi-organ disease caused by defects in ciliary structure and function. It results in a wide range of clinical manifestations, most commonly in the upper and lower airways. Central data collection in national and international registries is essential to studying the epidemiology of rare diseases and filling in gaps in knowledge of diseases such as PCD. For this reason, the Swiss Primary Ciliary Dyskinesia Registry (CH-PCD) was founded in 2013 as a collaborative project between epidemiologists and adult and paediatric pulmonologists.\n\nThe registry records patients of any age, suffering from PCD, who are treated and resident in Switzerland. It collects information from patients identified through physicians, diagnostic facilities, and patient organisations. The registry dataset contains data on diagnostic evaluations, lung function, microbiology and imaging, symptoms, treatments, and hospitalizations.\n\nBy May 2018, CH-PCD has contacted 566 physicians of different specialties and identified 134 patients with PCD. At present this number represents an overall 1 in 63,000 prevalence of people diagnosed with PCD in Switzerland. Prevalence differs by age and region; it is highest in children and adults younger than 30 years, and in Espace Mittelland. The median age of patients in the registry is 25 years (range 5-73), and 49 patients have a definite PCD diagnosis based on recent international guidelines. Data from CH-PCD are contributed to international collaborative studies and the registry facilitates patient identification for nested studies.\n\nCH-PCD has proven to be a valuable research tool that already has highlighted weaknesses in PCD clinical practice in Switzerland. Development of centralised diagnostic and management centres and adherence to international guidelines are needed to improve diagnosis and management--particularly for adult PCD patients.

epidemiology

Perfect Counterfactuals for Epidemic Simulations

Simulation studies are often used to predict the expected impact of control measures in infectious disease outbreaks. Typically, two independent sets of simulations are conducted, one with the intervetnion, and one without, and epidemic sizes (or some related metric) are compared to estimate the effect of the intervention. Since it is possible that controlled epidemics are larger than uncontrolled ones if there is substantial stochastic variation between epidemics, uncertainty intervals from this approach can include a negative effect even for an effective intervention. To more precisely estimate the number of cases an intervention will prevent within a single epidemic, here we develop a single world approach to matching simulations of controlled epidemics to their exact uncontrolled counterfac-tual. Our method borrows concepts from percolation approaches prune out possible epidemic histories and create potential epidemic graph that can be realized to create perfectly matched controlled and uncontrolled epidemics. We present an implementation of this method for a common class of compartmental models, and its application in a simple SIR model. Results illustrate how, at the cost of some computation time, this method substantially narrows confidence intervals and avoids non-sensical inferences.

epidemiology