bioRxiv Science⌕ Search

bioRxiv · 10.64898/2026.01.07.698260

Efficient in vitro refactoring and biosynthetic gene cluster amplification for the overproduction and accelerated discovery of anticancer thioamitides

Abstract

Thioamitides, a class of highly modified bacterial ribosomally synthesised and post-translationally modified peptides (RiPPs), have potent activities against multiple cancer cell lines. Among these compounds, the structurally divergent thioalbamide combines promising in vivo antiproliferative activity with a superior chemical stability respect to its counterparts. However, thioalbamide is produced in low yields by its genetically intractable native producer and its biosynthetic pathway was initially not productive when transferred into the heterologous host Streptomyces coelicolor M1146. These circumstances substantially hamper to increase the production of this promising compound. Here, we show how in vitro Gibson-like assemblies can be employed for the quick and efficient refactoring of the thioalbamide biosynthetic gene cluster (BGC), leading to substantially increased levels of production in S. coelicolor M1146 through a prioritised selection of promoters. Via this work, PtsrA and PgroEL2 were identified as beneficial additions to the Streptomyces synthetic biology toolbox. We then assessed bacterial genomes for biosynthetic gene clusters (BGCs) predicted to produce thioalbamide-like compounds with improved hydrophilicity. This rational discovery campaign led to the identification a silent thioamitide BGC encoding a thioalbamide-like core peptide but clustered with additional tailoring enzymes, including a previously unknown cupin-fold protein. Applying the refactoring strategy together with the simultaneous expression of multiple BGC copies, we characterised the product of this pathway, thiocupinamide, a polyhydroxylated thioamitide closely related to thioalbamide. We show that thiocupinamide has potent anticancer and antibacterial activities.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Santos-Aberturas, J., Frattaruolo, L., Tassinari, E., Rejzek, M., Fong, L. K. W., Kawicha, P., Sangdee, A., Grandellis, C., Cappello, A. R., Truman, A. W.. 2026-01-08. Efficient in vitro refactoring and biosynthetic gene cluster amplification for the overproduction and accelerated discovery of anticancer thioamitides. https://doi.org/10.64898/2026.01.07.698260

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Biocatalytic Production of Galantamine in Yeast through Cytochrome P450 Optimization

Galantamine is a pharmaceutically relevant Amaryllidaceae alkaloid used for the treatment of Alzheimer's disease. Its structural complexity and low abundance in plants motivate the development of alternative manufacturing routes. Here, we established the first engineered yeast platform for the biocatalytic production of galantamine from 4OMe-norbelladine. Heterologous expression of the downstream pathway enzymes NtCYP96T6, NtNMT1, and NtAKR1 from Narcissus cv. Tete-a-Tete in Saccharomyces cerevisiae was complemented by optimization of cultivation temperature, carbon source, and medium pH enabling the first demonstration of galantamine production in yeast. Systematic screening of cytochrome P450 reductase and cytochrome b partners to boost NtCYP96T6 activity led to the identification of a novel reductase mined from the Narcissus pseudonarcissus transcriptome, NpCPR, which supported the highest pathway flux and yielded 7.0 {+/-} 0.6 mg/L galantamine, corresponding to a 7.6 % molar yield from 250 M 4OMe-norbelladine and an approximately 173-fold improvement over the parental strain. We further exploited this yeast cell factory for the precursor-directed biosynthesis of 7F-galantamine, highlighting the potential of pathway enzyme promiscuity to access new-to-nature GAL analogues that may be challenging to produce through conventional chemical synthesis. Together, this work establishes a foundation for microbial galantamine production and biosynthetic diversification of its pharmaceutically relevant scaffold.

synthetic biology↗

Adaptive laboratory evolution of a yeast co-culture chassis for modular bioproduction

Synthetic microbial consortia have the potential to bring novel architectures to biotechnological processes. They enable more flexible and efficient bioprocesses by reducing metabolic burden and supporting division of labour. Obligately mutualistic cross-feeding has been used to stabilise the population composition of artificially assembled consortia. However, many current strategies rely on the cross-feeding of metabolites with limited exchange rates, and this imposes a growth burden on the consortium members. Here, we used adaptive laboratory evolution (ALE) to address this growth bottleneck and create an optimised co-culture chassis to host bioproduction functions. Transcriptome analysis revealed how ALE alleviated stress responses associated with nutritional restrictions in the non-evolved cross-feeding system. Using a case-study split bioproduction pathway, we demonstrated that improvements in the chassis growth performance resulted in production improvements, surpassing the performance of a monoculture implementation by 1.5-fold. Our results show the potential of ALE to optimise the cross-feeding layer of yeast co-cultures, and how this enables the efficient implementation of a bioproduction process.

synthetic biology↗

Generative Language Modeling for Antibody CDR Grafting and Alignment-driven De Novo Design

Antibodies recognise their targets through hypervariable complementarity-determining regions (CDRs), which are interleaved with conserved frameworks in sequence space, making de novo CDR design an infilling problem. Autoregressive models generate residues left-to-right, which precludes full framework context during CDR generation and conflates framework and CDR likelihoods, leaving no natural prompt-response interface for feedback to steer generation. We present GenCDR, a family of LLaMa-based autoregressive language models that read all frameworks as a conditioning prompt and generate all CDRs jointly as a variable-length response, making CDR likelihoods a clean, separable target for reward attribution. The family comprises IgGenCDR, p-IgGenCDR, and NanoGenCDR, trained on unpaired, paired, and nanobody chains, respectively. GenCDR achieves the highest CDR recovery among autoregressive models and produces natural, diverse, human-like CDRs whose likelihoods correlate with fitness and developability assays. The prompt-response boundary also enables principled alignment: reward signals for binding affinity, expression, or developability can be composed to steer CDR generation. Over four rounds of alignment against antibody-antigen co-folding and developability objectives, we find that NanoGenCDR, which uses no explicit antigen encoding, can reach in silico structural interface metrics competitive with those of a structure-conditioned diffusion pipeline at roughly half the sampling budget, with more natural, developable designs. The same interface can be extended to integrate experimental feedback, opening a path to closed-loop antibody de novo design.

synthetic biology↗