bioRxiv Science⌕ Search

bioRxiv · 10.1101/2023.05.12.540511

Giants among Cnidaria: large nuclear genomes and rearranged mitochondrial genomes in siphonophores

Abstract

Siphonophores (Cnidaria:Hydrozoa) are abundant predators found throughout the ocean and are important components in worldwide zooplankton. They range in length from a few centimeters to tens of meters. They are gelatinous, fragile, and difficult to collect, so many aspects of the biology of these 190 species remain poorly understood. To survey siphonophore genome diversity, we performed Illumina sequencing of 32 species sampled broadly across the phylogeny. Sequencing depth was sufficient to estimate nuclear genome size from k-mer spectra in 8 specimens, ranging from 0.7-4.8Gb. In 6 specimens we got heterozygosity estimates between 0.7-5.3%. Rarefaction analyses indicate k-mer peaks can be absent with as much as 30x read coverage, suggesting minimum genome sizes range from 1.0-3.8Gb in the remaining 27 samples without k-mer peaks. This work confirms most siphonophore nuclear genomes are large, but also identifies several with reduced size that are tractable targets for future siphonophore nuclear genome assembly projects. We also assembled mitochondrial genomes for 32 specimens from these new data, indicating a conserved gene order among Hydrozoa, Cystonectae and some Physonectae, also revealing the ancestral gene organization of siphonophores. There then was extensive rearrangement of mitochondrial genomes within other physonects and in Calycophorae, including the repeated loss of atp8. Though siphonophores comprise a small fraction of cnidarian species, this survey greatly expands our understanding of cnidarian genome diversity. This study further illustrates both the importance of deep phylogenetic sampling and the utility of Illumina genome skimming in understanding genomic diversity of a clade. SignificanceDescriptions of basic genome features, such as nuclear genome size and mitochondrial genome sequences, remain sparse across many clades in the tree of life, leading to over generalizations from very small sample sizes and often limiting selection of optimal species for genome assembly efforts. Here we use Illumina genome skimming to assess a variety of genome features across 35 siphonophores (Cnidaria). This deep dive within a single clade identifies six species that are optimal candidates of future genomic work, and reveals greater range in nuclear genome size and diversity of mitochondrial genome orders within siphonophores than had been described across all Cnidaria.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ahuja, N., Cao, X., Schultz, D. T., Picciani, N., Lord, A., Shao, S., Burdick, D. R., Haddock, S. H. D., Li, Y., Dunn, C. W.. 2023-05-14. Giants among Cnidaria: large nuclear genomes and rearranged mitochondrial genomes in siphonophores. https://doi.org/10.1101/2023.05.12.540511

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Denisovan introgression left differential selection regimes in Humans and Neanderthals on the SLC30A9 gene

Signals of positive selection around the SLC30A9 gene have been reported in human populations outside Africa. Selection likely acted on a highly differentiated single-nucleotide polymorphism, rs1047626, leading to a non-synonymous substitution in the encoded zinc transporter. Because of the striking similarity between the putatively selected SLC30A9 haplotype observed in several current human populations and the Denisovan individual, previous work has proposed adaptive introgression. Yet alternative explanations, including ancient human variation, and the precise archaic source -Neanderthal or Denisovan- remained unresolved. Considering the potentially complex evolution of SLC30A9, we applied Approximate Bayesian Computation (ABC) algorithms coupled to machine learning to investigate the most plausible evolutionary origin of this substitution. After modelling different evolutionary scenarios with forward-in-time simulations, our results highlight that the most probable scenario is a Denisovan origin of the rs1047626 polymorphism. However, the allele likely introgressed into Neanderthals first and was then passed into non-African modern humans. Moreover, the derived allele frequency for rs1047626 across several African populations is consistent with back-to-Africa migrations. Finally, our ABC analyses indicate strong positive selection in East Asian populations and other out-of-Africa populations, whereas in Neanderthal populations, the selection coefficient was probably neutral or slightly deleterious.

evolutionary biology↗

RELAX does not reproduce its own estimates at default settings, and its output does not show it

Selection-intensity estimates from RELAX are reported as a point value of K with a likelihood-ratio P. We report that, at default settings and on data of ordinary size, the program does not reproduce its own fits. Of 27 enzyme entries refitted under two optimiser configurations, none reproduced its log-likelihood to within 0.01 units; the median change was 103 units, the largest over 3,400, and four verdicts reversed. Eighty null orthologues reproduced none. A byte-identical command returned a distinct likelihood on every repetition, single-threaded, across three releases, and on alignments simulated under the fitted model, where 3.3 per cent of replicates reproduced. The documented random-number seed never reaches the generator when assigned on the command line, yet reads back as the value supplied. PAML localises the cause: its two-ratio model, without site classes, reproduced its log-likelihood for all 288 genes; its site-class models agreed for 27 to 67 per cent. The instability follows the mixture over sites, not the program. The output does not show it: 46 of 410 fits ended with a negative likelihood-ratio statistic, impossible under convergence, and 123 of 410 report a K re-estimated under a domain restriction rather than the unconstrained maximum. Of 234 published studies using RELAX, none reported a seed. Seeding while holding the thread count at one reproduced sixty of sixty runs on twenty genes under two releases; the seed alone reproduced none of five, and no documentation states the second condition. We recommend that fits be repeated and their dispersion published.

evolutionary biology↗

Sequential accumulation of adaptive alleles forms an inversion supergene in deer mice

Supergenes are clusters of co-inherited loci that affect multiple or complex phenotypes. Despite the growing number of chromosomal inversions identified as supergenes in natural populations, their molecular basis and evolutionary history often remain obscure. Here, we identified two candidate genes, Slc45a2 and Npr3, within a 41-Mb inversion supergene in the deer mouse (Peromyscus maniculatus) that respectively drive darker coats and longer tails - two traits associated with forest adaptation. Mice homozygous for the inversion (inv/inv) exhibit elevated Slc45a2 expression in melanocytes relative to the congenic standard genotype (std/std), disrupting pheomelanin production. In parallel, downregulation of Npr3 in inv/inv mouse growth plates prolongs postnatal growth of caudal vertebrae, resulting in tail elongation. Population-level analyses further implicate that this supergene arose through the subsequent accumulation of the Npr3 allele within the inversion, rather than by capturing all beneficial mutations at its origin.

evolutionary biology↗