bioRxiv Science⌕ Search

bioRxiv · 10.64898/2026.04.29.721750

Phylogenomic Taxonomic Analysis of Ralstonia solanacearum Strains causing Bacterial Wilt Disease in Northeastern Argentina.

Abstract

Ralstonia solanacearum species complex (RSSC) is a genetically diverse group of plant pathogens, yet genomic data from South America remain limited. Here, we characterize 13 RSSC strains isolated from tomato, pepper, and eggplant in northeastern Argentina. Phylogenetic analysis of the egl marker gene assigned these strains to phylotype IIA and suggested two closely related lineages. Complete genomes (5.63-5.76 Mb) were generated for four representative strains, yielding high-quality (99.94% completeness with f_Burkholderiaceae CheckM markers), closed assemblies with canonical bipartite architecture. Phylogenetic analysis of the egl marker, 49 conserved bacterial genes, and average nucleotide identity (ANI) analyses, consistently assigned one lineage to sequevar IIA-50, forming a coherent and monophyletic group. In contrast, although egl analysis suggested the second lineage was related to one sequevar IIA-38 reference strain, genomic analysis did not support this assignment. Further, the genomic analysis revealed significant genomic distance between the genomes for two sequevar 38 representative strains, supporting a conclusion that sequevar 38 itself was not monophyletic and instead appears paraphyletic. These findings highlight limitations of single-locus classification and support genome-informed refinement of RSSC sub-phylotype taxonomy. Outcome statementReports of bacterial wilt disease in Argentina had not yet been published in the international literature although the disease has been long-standing. This study provides complete genome sequences for four Ralstonia solanacearum strains from Northern Argentina and places them within a global phylogenomic framework. The Argentine strains cluster into two closely related phylotype IIA lineages, indicating that bacterial wilt in this regional dataset is associated with genetically similar populations. For clear communication of which strains are present in Northern Argentina, we attempted to classify the lineages to the long-standing sequence variant (sequevar) system for naming R. solanacearum species complex (RSSC) strains. One lineage was confidently assigned to IIA-50 with genomic support that confirmed phylogenetic analysis of the classical genetic marker egl. However, newly available genomes for sequevar reference strains revealed an issue where two distantly related strains are currently recognized as references for sequevars. Overall, these results provide evidence supporting the need for genome-informed refinement of sub-phylotype classification and expand genomic representation of South American RSSC populations. Data summaryComplete genome assemblies and raw reads for INTABV18, INTABV29, INTABV624 and INTABV2657 are deposited to NCBI under the project number PRJNA1407867. The curated dataset of public RSSC genomes is available to users who register a free account on KBase via a KBase narrative (https://narrative.kbase.us/narrative/189849). The narrative described in a living BioRxiv pre-print [1]. Supplemental files such as Figure S1, rectangular versions of all trees (Figure 2 and 3 and S1) and supplementary table S1, S2, S3 and S4 are available on Zenodo at doi.org/10.5281/zenodo.19502890 O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=172 SRC="FIGDIR/small/721750v1_figS1.gif" ALT="Figure 1"> View larger version (47K): org.highwire.dtl.DTLVardef@1ac3168org.highwire.dtl.DTLVardef@1dfd0d6org.highwire.dtl.DTLVardef@107ae42org.highwire.dtl.DTLVardef@141937c_HPS_FORMAT_FIGEXP M_FIG O_FLOATNOFigure S1.C_FLOATNO Maximum-likelihood phylogenetic tree inferred from 471 bp of endoglucanase (egl) gene sequences assigned Argentine strains as phylotype II sequevar 38 and sequevar 50. The tree was constructed using PhyML v3.0 under the GTR nucleotide substitution model with gamma-distributed rate heterogeneity ( = 0.33), as selected by the SMART model selection procedure implemented in PhyML (Lefort et al., 2017). The egl sequences from Argentine strains are highlighted in blue, and their corresponding GenBank accession numbers for both the egl nucleotide sequence and the whole-genome assembly are shown in parentheses. Reference egl sequences representing sequevars IIA-38 (CFBP6801 and CIP120) and IIA-50 (T1-UY and ACH1076) are also shown in bold and marked with yellow circles. A searchable PDF of this tree in rectangular format is available on Zenodo (doi.org/10.5281/zenodo.19502890). C_FIG O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=196 SRC="FIGDIR/small/721750v1_fig2.gif" ALT="Figure 2"> View larger version (53K): org.highwire.dtl.DTLVardef@39d776org.highwire.dtl.DTLVardef@170bd89org.highwire.dtl.DTLVardef@aba166org.highwire.dtl.DTLVardef@1f156dd_HPS_FORMAT_FIGEXP M_FIG O_FLOATNOFigure 2.C_FLOATNO Maximum-likelihood phylogenetic tree inferred from 710 bp of endoglucanase (egl) gene sequences assigned Argentine strains as phylotype II sequevar 38 and sequevar 50. The phylogenetic tree was constructed using PhyML v3.0 under the GTR+R nucleotide substitution model, as selected by the SMART model selection procedure (Lefort et al., 2017). egl sequences from four Argentine strains (INTABV18, INTABV29, INTABV624, and INTABV2657) are shown in bold and highlighted in blue. Reference egl sequences representing sequevars IIA-38 (CFBP6801 and CIP120) and IIA-50 (T1-UY and ACH1076) are also shown in bold and marked with yellow circles. Two USA strains identified as IIA-38 (UCD576 and RS124) are shown in bold. A searchable PDF of this tree in rectangular format is available on Zenodo (doi.org/10.5281/zenodo.19502890). C_FIG O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=116 SRC="FIGDIR/small/721750v1_fig3.gif" ALT="Figure 3"> View larger version (37K): org.highwire.dtl.DTLVardef@17dd372org.highwire.dtl.DTLVardef@1c5156corg.highwire.dtl.DTLVardef@179d9org.highwire.dtl.DTLVardef@e6d529_HPS_FORMAT_FIGEXP M_FIG O_FLOATNOFigure 3.C_FLOATNO Approximate maximum-likelihood phylogeny based on a concatenated alignment of 49 conserved genes places four Argentine genomes (INTABV18, INTABV29, INTABV624 and INTABV2657) within the phylotype IIA clade. The tree was constructed using the SpeciesTreeBuilder v0.1.4 application on the KBase platform, incorporating the four Argentine genomes into a reference dataset of 825 genomes representing the known global diversity of the RSSC. The tree was visualized and annotated using iTOL v7.4.2. Argentine genomes are shown in bold and highlighted in blue, and egl reference strains for the sequevar IIA-38 (CIP120 and CFBP6801) and IIA-50 (T1-UY) are shown in bold and marked with yellow circles. Branches with approximate likelihood-ratio support values higher than >70% are colored in blue. A searchable PDF of this tree in rectangular format is available on Zenodo (doi.org/10.5281/zenodo.19502890). C_FIG

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Obregon, V., Shin, G. Y., Galdeano, E., Escobar, R., Lattar, T., Ibanez, J. M., Amadio, A., Irazoqui, J. M., Santiago, G. M., Eberhardt, M. F., Gochez, A. M., Lowe-Power, T.. 2026-05-01. Phylogenomic Taxonomic Analysis of Ralstonia solanacearum Strains causing Bacterial Wilt Disease in Northeastern Argentina.. https://doi.org/10.64898/2026.04.29.721750

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Rapid and largely reversible shifts in the canine fecal metabolome during dietary change

Diet can rapidly change the fecal metabolome, but less is known about recovery after the original diet is restored. We used untargeted UPLC-MS metabolomics to analyze 72 fecal samples from nine Pumi dogs during an owner-managed switch from dry food to raw food and back to dry food. Diet phase accounted for a large proportion of variation in both ionization modes. More than 13,000 LC-MS features changed at the first sampling point after the switch to raw food, with a similarly large response after return to dry food. Among features significant in both comparisons, more than 99% changed in opposite directions. At the final sampling point, no positive-mode (ESI+) features and only 13 negative-mode (ESI-) features differed from the second dry-food baseline under the same threshold. BARF-associated patterns persisted in analyses excluding individual dogs and in pedigree-adjusted candidate models, although individual feature effects depended on normalization. Putative metabolites from several biochemical classes differed in their response and recovery. The fecal metabolome therefore changed rapidly and returned largely toward baseline, with differences among dogs.

microbiology↗

Taxonomic and functional concordance between full-length ONT 16S and ONT shotgun metagenomics in the canine gut microbiome

Background: Full-length Oxford Nanopore Technologies (ONT) 16S rRNA sequencing provides a scalable view of microbial community composition and can support phylogeny-based functional prediction, but it is not equivalent to shotgun metagenomics. We asked which biological conclusions are preserved when the same canine fecal specimens are profiled by full-length ONT 16S and ONT whole-genome shotgun (WGS) sequencing, and how their agreement depends on analytical scale, reference representation and classifier. Methods: Ninety-seven fecal specimens from 51 dogs were profiled with both assays from the same DNA extract. Functional profiles predicted from NanoASV/NanoPredict with PICRUSt2 were compared with WGS-supported KEGG Ortholog (KO) profiles generated by Kadath. Taxonomy was benchmarked in a source-genome-matched RefSeq universe and in a host-specific DogMAG universe using minitax and Kraken2. Agreement was evaluated at whole-profile, feature-abundance, detection, between-sample structure and biological-inference scales. Age-associated transfer was assessed with dog-aware continuous mixed models, grouped signed-score analyses and paired/dog-blocked PERMANOVA. Results: Functional whole-profile concordance was high: median within-sample CLR Spearman correlations ranged from 0.781 to 0.860 across developmental strata, while between-sample functional structure remained significant by Mantel (rho=0.543) and Procrustes (r=0.693; both p=0.001). Feature-wise transfer was substantially weaker (median KO-wise CLR Spearman=0.318). Continuous age-associated KO slopes showed substantial cross-assay concordance (Spearman=0.727; signed-score Spearman=0.753; direction agreement=77.9%), although 1,290/5,258 eligible KOs retained significant assay-by-age interactions. Taxonomically, exact genus/species abundance agreement was much lower than agreement in between-sample ecological structure. Host-specific DogMAG improved species-level median Spearman from 0.261 to 0.656 for minitax SpeciesEstimate and from 0.181 to 0.512 for Kraken2. The classifier effect was independent of reference choice: under both RefSeq and DogMAG, minitax yielded stronger 16S-WGS concordance than Kraken2, with all eight prespecified RefSeq paired genus/species endpoints and all 10 DogMAG primary paired endpoints significant after BH correction. The same ordering extended to developmental inference, with DogMAG genus/species age-slope concordance of 0.795/0.799 for SpeciesEstimate versus 0.693/0.702 for Kraken2. Taxonomic Aitchison PERMANOVA detected age-associated structure in every assay/reference/classifier/rank combination, whereas age-by-assay interactions were consistently significant but small (R2 approximately 1.1 to 2.2%). Stricter NanoASV identity thresholds removed substantial 16S abundance without improving species-level agreement. Conclusions: The extent of cross-assay agreement depends on the level of analysis. Full-length ONT 16S preserves broad functional organization, ecological structure and much of the direction of age-associated change, but exact fine-rank composition, individual-feature abundance and effect magnitude remain assay dependent. Host-specific reference representation substantially narrows the taxonomic gap, and classifier choice exerts an additional independent effect: within the same matched reference set, minitax consistently yields stronger 16S-WGS concordance than Kraken2 across abundance, detection, ecological-distance and developmental-inference endpoints. Full-length ONT 16S is therefore well suited to broad ecological screening and hypothesis generation, whereas WGS remains preferable when conclusions depend on quantitative fine-rank composition, directly supported gene content or precise feature-level effect estimates.

microbiology↗

Matrix-controlled emergence of biofilm architecture shapes antimicrobial survival

Biofilms are structured microbial communities whose extracellular matrix is widely regarded as a basis of their protection against antimicrobial compounds. Yet how matrix production by individual bacteria gives rise to collective architecture and antimicrobial protection remains poorly understood. Here, we systematically varied expression of the master biofilm regulator csgD in Salmonella enterica and found that increasing matrix production reorganizes biofilms from dense, isotropic packings into sparse, nematically aligned communities by altering cell-cell interactions. By combining experimentally measured biofilm architectures with reaction-diffusion modeling, we show that these structural changes produce distinct patterns of antimicrobial killing, ranging from preferential killing near the liquid-biofilm interface to more uniform killing throughout the community. Consequently, increasing matrix production unexpectedly reduces antimicrobial survival by shifting the biofilm into different transport regimes, while strain-specific physiological differences further modulate antimicrobial depletion. Rather than acting as a passive barrier, EPS therefore shapes antimicrobial susceptibility by reorganizing biofilm architecture and its transport properties. EPS thus provides a physical link between molecular regulation, collective architecture and antimicrobial survival, providing a quantitative framework for understanding how cellular matrix production generates emergent biofilm function.

microbiology↗