bioRxiv Science⌕ Search

bioRxiv · 10.1101/2024.07.01.601582

Hidden origami in Trypanosoma cruzi nuclei highlights its nonrandom 3D genomic organization

Abstract

The protozoan Trypanosoma cruzi, the causative agent of Chagas disease, exhibits polycistronic transcription and unidimensional genome compartmentalization of core (conserved) and disruptive (virulence factors from multigenic families) genes. Approximately 50% of its genome is repetitive, mainly virulence factor genes. Genomic sequences, including repeats, motifs of architectural proteins, and noncoding RNA loci are crucial for genome folding. Here, we evaluated the genomic features associated with higher-order chromatin organization in T. cruzi through extensive computational processing of high-throughput chromosome conformation capture (Hi-C) data, accounting for repetitive regions and improvements in genome annotation. Our study revealed that repetitive DNA (multimapped reads) influences 3D chromatin folding, particularly in determining the boundaries of topologically associated domains (TAD)-like structures. Virulence factor genes, unlike core genes, form shorter and more compact TAD-like structures enriched in loops, suggesting a gene expression regulatory mechanism. We found nonprotein-coding RNA loci (e.g., tRNAs) and transcription termination sites preferentially located at the boundaries of the TAD-like structures, while pseudogenes and multigenic family genes located in unstructured genomic regions. Our data indicate 3D clustering of tRNA loci, likely optimizing transcription by RNA polymerase III, and a complex interaction between spliced-leader RNA and 18S rRNA loci. Our findings provide insights into 3D genome organization in T. cruzi, contributing to the understanding of supranucleosome-level chromatin organization and suggesting possible links between 3D architecture and gene expression. We draw an analogy to the art of origami (e.g., papers folded into various shapes) resembling the DNA packed in chromatin fibers assuming distinct folds within the nucleus. ImportanceDespite the knowledge about the linear genome sequence and the identification of numerous virulence factors in the protozoan parasite Trypanosoma cruzi, there has been a limited understanding of how these genomic features are spatially organized within the nucleus and how this organization impacts gene regulation and pathogenicity. By providing a detailed analysis of the three-dimensional chromatin architecture in T. cruzi, our study contributed to filling this gap. We deciphered part of the origami structure hidden in the T. cruzi nucleus, showing the unidimensional genomic features are nonrandomly organized in the nuclear 3D landscape. We revealed the possible role of non-protein-coding RNA loci (e.g., tRNAs, SL-RNA, and 18S RNA) in shaping the genomic architecture. These findings provide insights into an additional epigenetic layer that may influence gene expression. Graphical abstractThe spatial organization of chromatin within the nuclei of T. cruzi and its resemblance to origami art. A. Identification of the 3D nuclear architectures within T. cruzi nuclei: topologically associating domains (TADs) and their boundaries; chromatin loops; and 3D networks. Inter- and intrachromosomal interactions reflect DNA-DNA contacts on the same (cis) and between different (trans) chromosomes. B. Resemblance between origami art and chromatin folding. Steps "a" to "l" show the process of folding a flat piece of paper from its unidimensional view up to its 3D boat form. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=150 SRC="FIGDIR/small/601582v1_ufig1.gif" ALT="Figure 1"> View larger version (48K): org.highwire.dtl.DTLVardef@1ca38b4org.highwire.dtl.DTLVardef@150bc99org.highwire.dtl.DTLVardef@18de8a9org.highwire.dtl.DTLVardef@1a5efa4_HPS_FORMAT_FIGEXP M_FIG C_FIG

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Bellini, N. K., de Lima, P. L. C., Pires, D. d. S., da Cunha, J. P. C.. 2024-07-05. Hidden origami in Trypanosoma cruzi nuclei highlights its nonrandom 3D genomic organization. https://doi.org/10.1101/2024.07.01.601582

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Sequence and epigenetic characterization of chromosome 21 centromeres in a family with recurrent Trisomy 21

Trisomy 21 (T21) is the most common genetic cause of intellectual disability, yet the molecular mechanisms underlying maternal meiosis I errors--responsible for ~70% of free T21 cases--remain poorly understood. In this preliminary study, we used long-read sequencing and genome assembly to investigate the DNA sequence and epigenetic features of chromosome 21 (chr21) centromeres in a family with recurrent free T21 due to maternal meiosis I errors. The mother, who had two affected and three unaffected children, showed no mosaicism or structural rearrangements. One of her two chr21 centromeres lacked a pronounced centromere dip region (CDR), displaying instead a diffuse hypomethylation pattern (dCDR) with much higher methylated CpG levels (55%) compared to its homologue (36%). This dCDR was transmitted to an unaffected child and the affected proband analyzed, suggesting it was present in one of the maternal chr21 since she was at least 32 years of age. Chr21 dCDRs were not observed in seven young mothers with children with T21 or previously described in the literature in 108 population haplotypes. We hypothesize that dCDRs may weaken kinetochore function, increasing nondisjunction risk, and propose two models linking such epigenetic variation to maternal age-related T21 risk. These findings highlight the value of complete centromere characterization in families with children with T21 and suggest centromere methylation status of chr21 as a potential T21 risk factor for future investigation.

genomics↗

Single-Cell Analytics for Dose Response (SCADR) discriminates PTEN missense variants by lipid and protein phosphatase dysfunction

The proliferation of sequencing efforts has revealed a vast and expanding catalog of single nucleotide gene variants, many associated to, but with unclear roles in disease. Fully charactering variant impacts and linking specific protein dysfunctions to disease are challenging due to the multi-functional nature of many proteins and varying degree of variant effects on these functions. Lagging are sensitive approaches to empirically assess the impact of missense variant-induced single amino acid changes on a wide range of protein functions. To address these issues, we have developed an open-source computational analysis tool called SCADR (Single-Cell Analytics for Dose Response) for simultaneously measuring and comparing impacts of exogenously-expressed variants on multiple signaling pathways using multiplex phospho-antibody spectral flow cytometry in human cell lines. SCADR retains and correlates single-cell measures of signal protein activity states along with expression levels of exogenously-expressed variants, providing rich characterization of multiple protein functions, signaling protein interactions, and enhanced discrimination of variant impacts on different signaling pathways, highlighting each variants unique dysfunction profile. Here, we apply SCADR for analyses of the impact of 6 variants of the tumor-suppressor protein PTEN (P38H, C124S, G129E, Y138L, D268E, 4A) expressed in HEK293 cells on the phosphorylation states of the canonical and noncanonical downstream signaling proteins Akt, S6, CREB, ERK, and p38 detected with fluorophore-conjugated phospho-antibodies, along with an antibody detecting an N-terminal HA tag on PTEN variants allowing measures of dose-response effects of each variants expression on signaling cascades. Results identify variant-specific impacts on downstream signaling cascades.

genomics↗

Microsecond molecular dynamics of SOD1 variants suggest a structural basis for divergent ALS clinical outcomes

Amyotrophic lateral sclerosis (ALS) is a fatal neurodegenerative disease characterised by progressive motor neuron degeneration. Mutations in the SOD1 gene represent the second most common genetic cause of ALS (ALS), and distinct SOD1 missense variants present with markedly different clinical profiles. A4V leads to an aggressive form of the disease (median survival [~]1y), H46R confers a mild, slowly progressive course and I113T exhibits an intermediate phenotype. The molecular basis by which these mutations produce divergent clinical outcomes remains poorly understood. We performed extensive classical molecular dynamics simulations of wild-type SOD1 and the three ALS-associated variants in the apo monomeric state to attempt to investigate the mechanisms behind such phenotypic differences. Structural stability, global compactness, and conformational flexibility, as well as analysis of collective motions between residues and estimation of free energy, were assessed. The H46R, A4V, and I113T variants exhibited distinct dynamic behaviours, highlighting differences in structural stability, local flexibility, and intramolecular interactions. These findings suggest that specific structural regions may contribute differently to protein dysfunction and could represent key elements for understanding the relationship between molecular dynamic properties and the differing clinical severity associated with these variants. Most strikingly, H46R exhibited exceptional structural stability across every analytical level, the lowest global deviation, most attenuated local flexibility, strongest internal dynamic coordination, and the deepest, most confined free energy basins of any system examined. This convergent multi-layered evidence of structural restraint provides a compelling mechanistic basis for the mild and slowly progressive clinical course of H46R ALS, suggesting that enhanced conformational rigidity, rather than bulk destabilisation, is the defining biophysical feature of this variant, and that its pathogenic mechanism operates through a route fundamentally decoupled from the aggregation-driven toxicity that characterises the more aggressive SOD1-ALS mutations.

genomics↗