bioRxiv Science⌕ Search

bioRxiv · 10.1101/2024.11.23.624631

The challenge of sequencing Chlamydia trachomatis and other bacterial STI genomes directly from clinical swabs: the optimum solution

Abstract

Rates of bacterial sexually transmitted infections (STIs) are rising and accessing their genomes provides information on strain evolution, circulating strains, and encoded antimicrobial resistance (AMR). Notable pathogens include Chlamydia trachomatis (CT), Neisseria gonorrhoeae (NG) and Treponema pallidum (TP), globally the most common bacterial STIs. Mycoplasma genitalium (MG) is also a bacterial STI which is of concern due to AMR development. These bacteria are also fastidious or hard to culture, and standard sampling methods lyse bacteria, completely preventing pathogen culture. Clinical samples contain large amounts of human and other microbiota DNA. These factors hinder the sequencing of bacterial STI genomes. We aimed to overcome these challenges in obtaining whole genome sequences, and evaluated four approaches using clinical samples from Argentina (39), Switzerland (14), and cultured samples from Finland (2) and Argentina (1). First, direct genome sequencing from swab samples was attempted through Illumina deep metagenomic sequencing, showing extremely low levels of target DNA, with under 0.01% of the sequenced reads being from the target pathogens. Second, host DNA depletion followed by Illumina sequencing was not found to produce enrichment in these very low load samples. Third, we tried a selective long-read approach with the new adaptive sequencing from Oxford Nanopore Technologies (ONT), which also did not improve enrichment sufficiently to provide genomic information. Finally, target enrichment using a novel pan-genome set of custom SureSelect probes targeting CT, NG, TP, and MG followed by Illumina sequencing was successful. We produced whole genomes from 64% of CT positive samples; from 36% of NG positive samples, and from 60% of TP positive samples. Additionally, we enriched MG DNA to gain partial genomes from 60% of samples. This is the first publication to date to utilize a pan-genome STI panel in target enrichment. Target enrichment, though costly, proved essential for obtaining genomic data from clinical samples. This data can be utilized to examine circulating strains, genotypic resistance, and guide public health strategies. Graphical Abstract O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=126 SRC="FIGDIR/small/624631v1_ufig1.gif" ALT="Figure 1"> View larger version (36K): org.highwire.dtl.DTLVardef@578aeborg.highwire.dtl.DTLVardef@16185f9org.highwire.dtl.DTLVardef@1a2ae87org.highwire.dtl.DTLVardef@1703fa7_HPS_FORMAT_FIGEXP M_FIG C_FIG Impact statementGenome data on circulating sexually transmitted infections (STIs) is important to better understand transmission networks, antimicrobial resistance and to guide treatment decisions. For many bacterial STIs, this information is difficult to obtain, as the bacteria are fastidious, in some cases intracellular, and often recalcitrant to culture. We have developed and tested a target enrichment STI panel of baits to capture whole genomes of Chlamydia trachomatis, Neisseria gonorrhoeae, Treponema pallidum, and Mycoplasma genitalium with approximately 50% success in genome sequencing for the first three pathogens. We compare this against other sequencing and enrichment methods, which did not provide sufficient data for genome analysis. This panel approach shows potential for clinical samples carrying these pathogens and can potentially also be developed for further pathogen groups. Data summaryAll illumina sequence data, with human read data removed using Hostile (1) and KrakenTools (https://github.com/jenniferlu717/KrakenTools), is deposited with the European Nucleotide Archive (ENA) under project number PRJEB72167.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Buttner, K. A., Bregy, V., Wegner, F., Purushothaman, S., Imkamp, F., Roloff Handschin, T., Puolakkainen, M. H., Hiltunen-Back, E., Braun, D., Kisakesen, I., Schreiber, A., Entrocassi, A. C., Gallo Vaulet, M. L., Lopez Aquino, D., Svidler Lopez, L., La Rosa, L., Egli, A., Rodriguez Fermepin, M., Seth-Smith, H.. 2024-11-23. The challenge of sequencing Chlamydia trachomatis and other bacterial STI genomes directly from clinical swabs: the optimum solution. https://doi.org/10.1101/2024.11.23.624631

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

A population-scale landscape of the subgingival microbiome reveals divergent routes to periodontal dysbiosis

Periodontitis is an archetypical mucosal inflammatory disease in which microbiome dysbiosis at the tooth-epithelial interface interacts with host genetic and behavioral risk factors to drive immune-mediated tissue destruction. Although subgingival microbiome compositional shifts are thought to parallel disease severity, microbiome variation at the population-level and its relationship to periodontal clinical phenotypes and disease-modifying factors remain poorly defined. Here, we use unsupervised manifold learning to map the compositional landscape of the subgingival microbiome in 1,355 adults spanning periodontal health to severe periodontitis. We identified eight latent microbiome states organized along a branching continuum from eubiosis to dysbiosis. An intermediate microbial configuration marked ecological destabilization and bifurcation into two distinct periodontitis-associated dysbiotic trajectories, distinguished by links to gingival inflammation and smoking. Although the microbiome trajectories broadly tracked periodontal destruction, a minority of individuals showed discordant microbiome-clinical phenotypes, with some individuals with periodontitis retaining otherwise eubiotic microbiomes enriched for low-abundance pathobionts, while some cases of health or mild disease had highly dysbiotic communities, suggesting distinct host susceptibility. Together, these findings define a population-scale ecological landscape of the subgingival microbiome, reveal divergent trajectories to periodontal dysbiosis, and highlight heterogeneity in the relationship between microbial community structure and clinical disease expression.

microbiology↗

Rapid and largely reversible shifts in the canine fecal metabolome during dietary change

Diet can rapidly change the fecal metabolome, but less is known about recovery after the original diet is restored. We used untargeted UPLC-MS metabolomics to analyze 72 fecal samples from nine Pumi dogs during an owner-managed switch from dry food to raw food and back to dry food. Diet phase accounted for a large proportion of variation in both ionization modes. More than 13,000 LC-MS features changed at the first sampling point after the switch to raw food, with a similarly large response after return to dry food. Among features significant in both comparisons, more than 99% changed in opposite directions. At the final sampling point, no positive-mode (ESI+) features and only 13 negative-mode (ESI-) features differed from the second dry-food baseline under the same threshold. BARF-associated patterns persisted in analyses excluding individual dogs and in pedigree-adjusted candidate models, although individual feature effects depended on normalization. Putative metabolites from several biochemical classes differed in their response and recovery. The fecal metabolome therefore changed rapidly and returned largely toward baseline, with differences among dogs.

microbiology↗

Taxonomic and functional concordance between full-length ONT 16S and ONT shotgun metagenomics in the canine gut microbiome

Background: Full-length Oxford Nanopore Technologies (ONT) 16S rRNA sequencing provides a scalable view of microbial community composition and can support phylogeny-based functional prediction, but it is not equivalent to shotgun metagenomics. We asked which biological conclusions are preserved when the same canine fecal specimens are profiled by full-length ONT 16S and ONT whole-genome shotgun (WGS) sequencing, and how their agreement depends on analytical scale, reference representation and classifier. Methods: Ninety-seven fecal specimens from 51 dogs were profiled with both assays from the same DNA extract. Functional profiles predicted from NanoASV/NanoPredict with PICRUSt2 were compared with WGS-supported KEGG Ortholog (KO) profiles generated by Kadath. Taxonomy was benchmarked in a source-genome-matched RefSeq universe and in a host-specific DogMAG universe using minitax and Kraken2. Agreement was evaluated at whole-profile, feature-abundance, detection, between-sample structure and biological-inference scales. Age-associated transfer was assessed with dog-aware continuous mixed models, grouped signed-score analyses and paired/dog-blocked PERMANOVA. Results: Functional whole-profile concordance was high: median within-sample CLR Spearman correlations ranged from 0.781 to 0.860 across developmental strata, while between-sample functional structure remained significant by Mantel (rho=0.543) and Procrustes (r=0.693; both p=0.001). Feature-wise transfer was substantially weaker (median KO-wise CLR Spearman=0.318). Continuous age-associated KO slopes showed substantial cross-assay concordance (Spearman=0.727; signed-score Spearman=0.753; direction agreement=77.9%), although 1,290/5,258 eligible KOs retained significant assay-by-age interactions. Taxonomically, exact genus/species abundance agreement was much lower than agreement in between-sample ecological structure. Host-specific DogMAG improved species-level median Spearman from 0.261 to 0.656 for minitax SpeciesEstimate and from 0.181 to 0.512 for Kraken2. The classifier effect was independent of reference choice: under both RefSeq and DogMAG, minitax yielded stronger 16S-WGS concordance than Kraken2, with all eight prespecified RefSeq paired genus/species endpoints and all 10 DogMAG primary paired endpoints significant after BH correction. The same ordering extended to developmental inference, with DogMAG genus/species age-slope concordance of 0.795/0.799 for SpeciesEstimate versus 0.693/0.702 for Kraken2. Taxonomic Aitchison PERMANOVA detected age-associated structure in every assay/reference/classifier/rank combination, whereas age-by-assay interactions were consistently significant but small (R2 approximately 1.1 to 2.2%). Stricter NanoASV identity thresholds removed substantial 16S abundance without improving species-level agreement. Conclusions: The extent of cross-assay agreement depends on the level of analysis. Full-length ONT 16S preserves broad functional organization, ecological structure and much of the direction of age-associated change, but exact fine-rank composition, individual-feature abundance and effect magnitude remain assay dependent. Host-specific reference representation substantially narrows the taxonomic gap, and classifier choice exerts an additional independent effect: within the same matched reference set, minitax consistently yields stronger 16S-WGS concordance than Kraken2 across abundance, detection, ecological-distance and developmental-inference endpoints. Full-length ONT 16S is therefore well suited to broad ecological screening and hypothesis generation, whereas WGS remains preferable when conclusions depend on quantitative fine-rank composition, directly supported gene content or precise feature-level effect estimates.

microbiology↗