bioRxiv Science⌕ Search

bioRxiv · 10.64898/2026.06.02.729587

Multi-omic data fusion reveals the in vivo enzyme kinetics of Vibrio natriegens at the genome-scale

Abstract

Vibrio natriegens is a halophilic, Gram-negative marine bacterium that is increasingly used in metabolic engineering applications due to its fast growth rate. In sparse minimal medium the organism has a doubling time of 25 minutes, which is about twice as fast as Escherichia coli under similar conditions. Given that its protein density is similarly constrained to that of E. coli, this necessitates that its metabolic enzymes are able to catalyze flux at a higher rate to sustain its metabolism. In this work, we measure the apparent turnover numbers of metabolically active enzymes in V. natriegens under a variety of growth conditions. The apparent turnover numbers of V. natriegens enzymes were measured in vivo by conducting coupled quantitative proteomics and 13C metabolic flux analysis experiments under seven different carbon source conditions in sparse minimal medium. A high quality genome-scale metabolic model was constructed and curated using additional experimental data. This model was extended with enzyme constraints, and subsequently used to find kinetic parameters that minimize the difference between model predictions and experimental observations. This model guided data fusion approach enabled the estimation of 357 apparent turnover numbers for metabolically active enzymes in V. natriegens. Our results reveal that the metabolic enzymes of V. natriegens are in median 14-fold faster than those of E. coli under similar conditions. Moreover, we show that machine learning generated turnover number estimates substantially underestimate the kinetics of V. natriegens. Our turnover number estimates were used to parameterize multiple condition dependent enzyme constrained flux balance analysis models of V. natriegens, which improved their predictive accuracy compared to the machine learning parameterisation. The combined experimental-computational approach employed here sheds light on the mechanism V. natriegens uses to accelerate its growth. This approach can also be extended to other bacteria, increasing the availability of in vivo measured enzyme turnover numbers, and improving the predictive accuracy of enzyme constrained metabolic models of other microbes.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Wilken, S. E., Beyss, M., Kratochvil, M., Grebel, A., Methling, K., Stefanski, A., London, P., Lalk, M., Schaper, K., Axmann, I. M., Noeh, K. M., Westhoff, P., Ebenhoeh, O.. 2026-06-04. Multi-omic data fusion reveals the in vivo enzyme kinetics of Vibrio natriegens at the genome-scale. https://doi.org/10.64898/2026.06.02.729587

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

SpaReg: sparsity-based 3D reconstruction of tissue microenvironments at native resolution across morphological and spatial molecular modalities

Tissue microenvironments comprise cellular and acellular components whose three-dimensional (3D) architecture guides disease fate. Direct imaging of intact specimens by light-sheet and multiphoton microscopy, and computational reconstruction from serial sections, have established that 3D spatial context reveals cell and tissue organization inaccessible at single planes. Computational reconstruction in particular can leverage archived human tissue, benefiting from the cost-effectiveness, robustness, scalable storage, workflow compatibility, and century-long pathobiology knowledge of histology, and can integrate multiple spatial modalities. However, sectioning can introduce tears and folds, and computational alignment can further distort tissue integrity. Here we introduce SpaReg, a sparsity-based 3D reconstruction method spanning histology, spatial proteomics and spatial transcriptomics. Across multiple organs, SpaReg robustly reconstructs large tissue volumes with preserved subcellular morphology despite sectioning artifacts. On a standardized histology benchmark, SpaReg achieves the best balance between 3D reconstruction accuracy and tissue integrity, and on spatial transcriptomics benchmarks it ranks among the leading methods while scaling to hundreds of sections and millions of cells in a dataset that several existing methods fail to process. Preservation of subcellular morphology by SpaReg also enables training of a Hematoxylin and Eosin (H&E)-based epithelial, T and B cell classifier, generating single-cell-resolved 3D maps directly from H&E. Applied to pancreatic tissue containing pancreatic ductal adenocarcinoma arising from an intraductal papillary mucinous neoplasm, these maps reveal that 2D sections overestimate immune exclusion, and resolve lymphoid aggregates in 3D. SpaReg, therefore, provides a scalable foundation for morphologically faithful, multimodal 3D atlases and spatially informed disease modeling

systems biology↗

TxCyto: A machine learning framework for estimating cytokine activity from whole transcriptome

Cytokines are critical mediators of intercellular communication, and a comprehensive characterization of their activity is essential for understanding health and disease. Existing tools to infer cytokine activity rely on experimental measurements. However, such measurements are available only for a small minority (43) of cytokines, and moreover, cytokine activity and response are highly context-specific, making a comprehensive experimental profiling across tissues, disease states, and biological contexts impractical. To address this gap, we developed TxCyto - a deep learning-based framework that infers the activity of cytokines, and more broadly of the tumor secretome, directly from the whole transcriptome profile of a sample. Trained on pan-cancer TCGA tumor transcriptomes, TxCyto was extensively validated in multiple independent datasets, including cytokine perturbation experiments. Across multiple cancer immunotherapy cohorts, TxCyto identified cytokines whose predicted activity was associated with therapeutic response. Furthermore, in spatial transcriptomic data for Liver cancer, TxCyto discovered spatial niches associated with response to immunotherapy. Overall, we develop a machine learning tool -TxCyto, for predicting the activity of 645 cytokines and tumor secretome from readily available whole transcriptomes. The TxCyto framework is generally applicable to other classes of regulatory molecules and TxCyto code base, and the tools are provided at https://github.com/Rahulncbs/TxCyto.

systems biology↗

Interpretable machine learning coupled to gene regulatory networks uncovers subcircuits underlying cell fate decisions

Gene regulatory networks (GRNs) model causal linkages that control cell fate decisions and differentiation transitions. Prioritizing regulatory subnetworks underlying cell state differences is of critical importance, but current methods including those reliant on topological metrics introduce circularity as the metrics prioritizing TFs are computed from the same networks whose assumptions they inherit. Separately, interpretable machine learning methods can identify latent factors (LFs) that discriminate cellular states with formal statistical guarantees but do not model regulatory linkages. Here, we present FOCAL (Factor-Outcome Coupling for Assessment of Linkages), a paradigm to prioritize regulatory subnetworks by coupling state-specific and dynamic GRNs with outcome-supervised LFs learned using interpretable machine learning without reference to network topology. This shifts GRN focus from macroscopic TF nodes to state-specific and dynamic TF-gene linkages. In B and T cells, FOCAL identified GIFs (GRNs coupled to Interpretable latent Factors), prioritized regulatory subnetworks underlying established states as well as transient regulatory episodes preceding them. By coupling LFs learnt from perturbation experiments of lineage-defining TFs, FOCAL identified transcriptional predisposition to alternative fates within progenitor cell populations before overt differentiation. This uncovered a novel NFATC2-IRF8 interplay in activated B cells, that was validated by in-vitro and in-vivo genetic perturbations. The two transcription factors act cooperatively to restrain extrafollicular plasmablast differentiation and promote germinal center B cell fate.

systems biology↗