bioRxiv Science⌕ Search

bioRxiv · 10.1101/2025.03.19.644114

Nonenzymatic RNA copying with a potentially primordial genetic alphabet

Abstract

Nonenzymatic RNA copying is thought to have been responsible for the replication of genetic information during the origin of life. However, chemical copying with the canonical nucleotides (A, U, G, and C) strongly favors the incorporation of G and C and disfavors the incorporation of A and especially U, because of the stronger G:C vs. A:U base pair, and the weaker stacking interactions of U. Recent advances in prebiotic chemistry suggest that the 2-thiopyrimidines were precursors to the canonical pyrimidines, raising the possibility that they may have played an important early role in RNA copying chemistry. Furthermore, 2-thiouridine (s2U) and inosine (I) form by deamination of 2-thiocytidine (s2C) and A respectively. We used thermodynamic and crystallographic analyses to compare the I:s2C and A:s2U base pairs. We find that the I:s2C base pair is isomorphic and isoenergetic with the A:s2U base pair. The I:s2C base pair is weaker than a canonical G:C base pair, while the A:s2U base pair is stronger than the canonical A:U base pair, so that a genetic alphabet consisting of s2U, s2C, I and A generates RNA duplexes with uniform base pairing energies. Consistent with these results, kinetic analysis of nonenzymatic template-directed primer extension reactions reveals that s2C and s2U substrates bind similarly to I and A in the template, and vice versa. Our work supports the plausibility of a potentially primordial genetic alphabet consisting of s2U, s2C, I and A, and offers a potential solution to the long-standing problem of biased nucleotide incorporation during nonenzymatic template copying. Significance StatementA long-standing challenge in primordial nonenzymatic RNA copying chemistry is the biased incorporation of C and G over A and U due to differences in base pair strength. We hypothesized that 2-thiopyrimidine substitution could help overcome this bias since A:s2U is a stronger version of the A:U base pair, and I:s2C is a weaker version of the G:C base pair. This study explores the efficacy of a potentially primordial genetic alphabet consisting of s2U, s2C, A and I. Our results show that A:s2U and I:s2C pairs are isoenergetic and isomorphic. Our findings highlight the potential of this alternative genetic alphabet to yield a more balanced incorporation of all nucleotides, facilitating information propagation by nonenzymatic RNA copying during the origin of life.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Fang, Z., Jia, X., Xing, Y., Szostak, J. W.. 2025-03-19. Nonenzymatic RNA copying with a potentially primordial genetic alphabet. https://doi.org/10.1101/2025.03.19.644114

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Conjunctive Targeting Links Drug Synergy to Emergent Proteome Structural States

Combinatorial therapies are widely used in the treatment of acute myeloid leukemia (AML) to address disease heterogeneity, adaptive resistance, and rewired signaling and metabolic states. Yet drug prioritization remains largely guided by clinical or phenotypic evidence, while the molecular mechanisms underlying effective drug combinations remain incompletely defined. To narrow this gap, we developed Combinatorial high-ratio Partial proteolysis with reference PRoteome Analysis (CoPPRA), a structural proteomics workflow based on limited proteolysis of cell lysates that profiles drug-associated changes in regional protein accessibility at peptide-level resolution. Here, we applied CoPPRA to ruxolitinib and ulixertinib, individually and in combination, in AML-related cell lysates. Our findings extend conjunctive targeting (CT), a recently proposed mechanism of combinatorial drug action in which combined exposure produces protein targeting patterns not observed with either drug alone. Previously identified through combination-associated changes in protein solubility/stability, CT is examined here at peptide-level resolution through regional differences in proteolytic accessibility. The ruxolitinib-ulixertinib combination produced broad peptide-level accessibility changes, including a subset meeting the predefined criteria for CT. CT candidates predominantly exhibited regional accessibility changes, with altered peptide regions occurring against comparatively small changes across the remaining quantified peptides from the same proteins. MAP2K1 and ATP6V1G1 showed pronounced differences between overlapping peptide sequences, highlighting localized variation in combination-associated accessibility, including an ATP6V1G1 peptide mapping to an annotated helical region. Combination-associated increases in peptide signals were also observed in PIK3R1, BRD4, and PTPN11, linking regional accessibility changes to signaling and transcriptional regulators relevant to AML. Functional enrichment and network analyses further implicated nucleotide and glucose metabolism, ficolin-1-rich granules, ribosome-associated processes, and phagocytic vesicles. These results extend conjunctive targeting from protein-level solubility/stability changes to regional differences in proteolytic accessibility, showing that combination-associated effects can be concentrated within specific peptide regions rather than distributed uniformly across proteins. More broadly, CoPPRA provides a peptide-resolved approach for investigating the molecular features of combinatorial drug action and prioritizing protein regions for subsequent mechanistic validation.

biochemistry↗

Structural and biochemical characterisation of an iterative GCN5-related N-acetyltransferase required for fungal siderophore tailoring

Siderophore-mediated iron acquisition is essential for fungal survival, particularly under iron-limiting conditions. In Aspergillus fumigatus, SidG, a member of the GCN5-related N-acetyltransferase (GNAT) superfamily, catalyses the final step in the biosynthesis of the extracellular siderophore triacetylfusarinine C (TAFC) through sequential acetylation of the precursor fusarinine C (FsC). However, the timing, catalytic mechanism, and functional significance of this modification are not fully understood. Here, we reconstituted SidG activity in vitro and combined native mass spectrometry, X-ray crystallography, molecular dynamics simulations, and site-directed mutagenesis to investigate its catalytic properties. Our analyses demonstrate that SidG selectively binds acetyl-CoA from the cellular milieu and iteratively acetylates the FsC scaffold prior to iron chelation. Structural, biochemical, and molecular dynamics analyses support a direct transfer mechanism, identify key catalytic residues, and demonstrate the strict selectivity of SidG for short-chain acyl-CoA donors. Together, these findings establish the molecular basis for SidG-dependent siderophore tailoring and expand our understanding of GNAT-catalysed transformations in fungal natural product biosynthesis.

biochemistry↗

Reconstitution of +1 nucleosome transcription reveals coordinated functions of SAGA, Mediator, and TFIIH

The +1 nucleosome has emerged as a key regulator of eukaryotic transcription, but how it controls transcription initiation remains poorly understood. Here we reconstitute transcription through the +1 nucleosome using eleven purified yeast factors: RNA polymerase II (Pol II), the six general transcription factors (GTFs), TFIIS, the activator Pho4, and the SAGA and Mediator complexes. The system recapitulates key features of regulation observed in vivo. SAGA, acting with Pho4, directs pre-initiation complex (PIC) assembly to the correct position through its TBP-loading activity. Mediator stimulates transcription when the +1 nucleosome imposes a barrier to PIC formation, consistent with stabilization of productive TFIIH-DNA engagement. Contrary to the prevailing model, SAGA remains bound to the PIC after TBP loading and acetylates the +1 nucleosome within the assembled complex. The isolated PIC-Mediator-SAGA-nucleosome complex is transcriptionally active, and the repressive effect of the nucleosome is relieved by the DNA translocase activity of Ssl2, the TFIIH subunit that opens promoter DNA. TFIIH thus couples promoter melting to remodeling of the +1 nucleosome.

biochemistry↗