bioRxiv Science⌕ Search

bioRxiv · 10.1101/775932

Virulence and antibiotic resistance plasticity of Arcobacter butzleri: insights on the genomic diversity of an emerging human pathogen

Abstract

Arcobacter butzleri is a food and waterborne bacteria and an emerging human pathogen, frequently displaying a multidrug resistant character. Still, no comprehensive genome-scale comparative analysis has been performed so far, which has limited our knowledge on A. butzleri diversification and pathogenicity. Here, we performed a deep genome analysis of A. butzleri focused on decoding its core- and pan-genome diversity and specific genetic traits underlying its pathogenic potential and diverse ecology. In total, 49 A. butzleri strains (collected from human, animal, food and environmental sources) were screened.\n\nA. butzleri (genome size 2.07-2.58 Mbp) revealed a large open pan-genome with 7474 genes (about 50% being singletons) and a small core-genome with 1165 genes. The core-genome is highly diverse ([≥]55% of the core genes presenting at least 40/49 alleles), being enriched with genes associated with housekeeping functions. In contrast, the accessory genome presented a high proportion of loci with an unknown function, also being particularly overrepresented by genes associated with defence mechanisms. A. butzleri revealed a plastic virulome (including newly identified determinants), marked by the differential presence of multiple adaptation-related virulence factors, such as the urease cluster ureD(AB)CEFG (phenotypically confirmed), the hypervariable hemagglutinin-encoding hecA, a putative type I secretion system (T1SS) harboring another agglutinin potentially related to adherence and a novel VirB/D4 T4SS likely linked to interbacterial competition and cytotoxicity. In addition, A. butzleri harbors a large repertoire of efflux pumps (EPs) (ten \"core\" and nine differentially present) and other antibiotic resistant determinants. We provide the first description of a genetic determinant of macrolides resistance in A. butzleri, by associating the inactivation of a TetR repressor (likely regulating an EP) with erythromycin resistance. Fluoroquinolones resistance correlated with the Thr-85-Ile substitution in GyrA and ampicillin resistance was linked to an OXA-15-like {beta}-lactamase. Remarkably, by decoding the polymorphism pattern of the porin- and adhesin-encoding main antigen PorA, this study strongly supports that this pathogen is able to exchange porA as a whole and/or hypervariable epitope-encoding regions separately, leading to a multitude of chimeric PorA presentations that can impact pathogen-host interaction during infection. Ultimately, our unprecedented screening of short sequence repeats detected potential phase-variable genes related to adaptation and host/environment interaction, such as lipopolysaccharide modification and motility/chemotaxis, suggesting that phase variation likely modulate A. butzleri key adaptive functions.\n\nIn summary, this study constitutes a turning point on A. butzleri comparative genomics revealing that this human gastrointestinal pathogen is equipped with vast virulence and antibiotic resistance arsenals, which, coupled with its remarkable core- and pan-genome diversity, opens a multitude of phenotypic fingerprints for environmental/host adaptation and pathogenicity.\n\nIMPACT STATEMENTDiarrhoeal diseases are the most common cause of human illness caused by foodborne hazards, but the surveillance of diarrhoeal diseases is biased towards the most commonly searched infectious agents (namely Campylobacter jejuni and C. coli). In fact, other less studied pathogens are frequently found as the etiological agent when refined non-selective culture conditions are applied. A hallmark example is the diarrhoeal-causing Arcobacter butzleri which, despite being also associated with extra-intestinal diseases, such as bacteremia in humans and mastitis in animals, and displaying high rates of antibiotic resistance, has not yet been profoundly investigated regarding its epidemiology, diversity and pathogenicity. To overcome the general lack of knowledge on A. butzleri comparative genomics, we provide the first comprehensive genome-scale analysis of A. butzleri focused on exploring the intraspecies virulome content and diversity, resistance determinants, as well as how this pathogen shapes its genome towards ecological adaptation and host invasion. The unveiled scenario of A. butzleri rampant diversity and plasticity reinforces the pathogenic potential of this food and waterborne hazard, while opening multiple research lines that will certainly contribute to the future development of more robust species-oriented diagnostics and molecular surveillance of A. butzleri.\n\nDATA SUMMARYA. butzleri raw sequence reads generated in the present study were deposited in the European Nucleotide Archive (ENA) (BioProject PRJEB34441). The assembled contigs (.fasta and .gbk files), the nucleotide sequences of the predicted transcripts (CDS, rRNA, tRNA, tmRNA, misc_RNA) (.ffn files) and the respective amino acid sequences of the translated CDS sequences (.faa files) are available at http://doi.org/10.5281/zenodo.3434222. Detailed ENA accession numbers, as well as the draft genome statistics are described in Table S1.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Isidro, J., Ferreira, S., Pinto, M., Domingues, F., Oleastro, M., Gomes, J. P., Borges, V.. 2019-09-19. Virulence and antibiotic resistance plasticity of Arcobacter butzleri: insights on the genomic diversity of an emerging human pathogen. https://doi.org/10.1101/775932

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Genomic correlates of metastatic competence and progression in human melanoma

Genomic events and their timing that grant a primary tumour the competence to disseminate remain poorly defined. We performed sequencing of 247 stage I/II primary cutaneous melanomas (CMs) and 60 matched metastases without intervening therapy from a prospectively followed registry cohort with a median followup of 92 months, integrating copy-number, mutational, protein and spatial-transcriptomic analyses. Relapse was not distinguished by oncogenic point mutations, which were largely shared between primaries and metastases, but by somatic copy-number alterations (SCNAs) and global chromosomal instability. We defined OncoCycle, a six-gene copy-number signature (amplification of CDK4, MCL1 and CD276; biallelic loss of CDKN2A, CDKN2B and TP53BP1) that predicted relapse independently of established clinicopathological features in melanoma, and a pan-cancer analysis. In matched pairs, metastatic progression was driven by continued copy-number evolution and reduction in intra-tumoural heterogeneity, rather than by acquired point mutations, and OncoCycle alterations from primary tumours were preserved in metastasis seeding clones. Clonal reconstruction revealed both monoclonal and polyclonal metastasis seeding, and spatial transcriptomics resolved copy-number-defined metastatic subclones occupying and programming distinct immune and stromal niches. Thus, metastatic competence was primed early by focal SCNAs on a background of chromosomal instability, elaborated by continued copy-number evolution during dissemination and spatio-temporal interactions with the tumour-microenvironment.

genomics↗

Identifying, phasing, and structurally annotating sex chromosomes for genome assemblies using CBS-tools

A complete reference genome for species with chromosomally-determined separate sexes should contain scaffolds for all sex chromosome homologs. However, sex chromosomes present distinct computational challenges compared to autosomes. Here we present a k-mer based analysis that utilizes whole-genome sequencing of a few sex-identified isolates: Cytogenetics-By-Sequencing (CBS) tools. Unlike other approaches that typically address one aspect of the sex chromosomes, CBS-tools strives to guide users from the discovery of the heterogametic sex through identifying the sex-determination region (SDR). The core of CBS-tools is automated quantification of sex-specific k-mers in order to predict the heterogametic sex. Using publicly-available datasets, CBS-tools correctly identified the known sex-system of the 31 species tested. Additionally, we used these k-mers to verify and correct phasing of sex chromosomes between haplotypes in species representing different sex-systems. Finally, we used these k-mers to delimit the SDR boundary using an interactive web platform. CBS-tools was developed with previously unexplored sex chromosome systems in mind, but is also suitable for well-examined sex chromosome pairs.

genomics↗

Evolutionary dynamics of the insertion sequence IS6110 in the Mycobacterium tuberculosis complex

Insertion sequences (IS) are the most common type of transposable element in prokaryotes and shape the structure of genomes through transposition and by providing a substrate for recombination. Despite the mutational impact of IS, the evolutionary dynamics of most elements in host species remain unknown. Here we study the dynamics of IS6110 in 10,000 strains of the Mycobacterium tuberculosis complex (MTBC). We developed a tool that allows the detection and comparison of IS insertions from short reads without using a reference genome. Using ancestral state reconstruction (ASR) on presence-absence patterns of IS6110, we describe the distribution of copy numbers (CNs) in the MTBC, infer birth rates of the element, and identify genomic regions with large numbers of parallel IS6110 insertions. Copy numbers in the MTBC range from 1 in some clades to more than 30 in strains of La3 (M. orygis). IS6110 birth rates scale approximately linearly with copy number and are elevated on terminal branches, consistent with the delayed action of purifying selection. A key characteristic of IS6110 is its occurrence in hotspots: the 5% most frequently targeted regions account for half of all independent insertion events. The motif 5'-TCTCAAAW-3' is enriched around target sites and in hotspots, suggesting that the accumulation of insertions in these regions results through a combination of non-random insertion and purifying selection in other regions. To conclude the study, we propose a niche constraints model according to which the distribution of IS6110 in the MTBC is governed by the rarity of regions that have both suitable DNA properties and little functional value for the host.

genomics↗