bioRxiv Science⌕ Search

bioRxiv · 10.1101/2022.10.04.510777

Identifying Essential Genes in Genome-Scale Metabolic Models of Consensus Molecular Subtypes of Colorectal Cancer

Abstract

Identifying essential targets in genome-scale metabolic networks of cancer cells is a time-consuming process. This study proposed a fuzzy hierarchical optimization framework for identifying essential genes, metabolites and reactions. On the basis of four objectives, the framework can identify essential targets that lead to cancer cell death, and evaluate metabolic flux perturbations of normal cells due to treatment. Through fuzzy set theory, a multiobjective optimization problem was converted into a trilevel maximizing decision-making (MDM) problem. We applied nested hybrid differential evolution to solve the trilevel MDM problem to identify essential targets in the genome-scale metabolic models of five consensus molecular subtypes (CMSs) of colorectal cancers. We used various media to identify essential targets for each CMS, and discovered that most targets affected all five CMSs and that some genes belonged to a CMS-specific model. We used the experimental data for the lethality of cancer cell lines from the DepMap database to validate the identified essential genes. The results reveal that most of the identified essential genes were compatible to colorectal cancer cell lines from DepMap and that these genes could engender a high percentage of cell death when knocked out, except for EBP, LSS and SLC7A6. The identified essential genes were mostly involved in cholesterol biosynthesis, nucleotide metabolisms, and the glycerophospholipid biosynthetic pathway. The genes in the cholesterol biosynthetic pathway were also revealed to be determinable, if the medium used excluded a cholesterol uptake reaction. By contrast, the genes in the cholesterol biosynthetic pathway were non-essential, if a cholesterol uptake reaction was involved in the medium used. Furthermore, the essential gene CRLS1 was revealed as a medium-independent target for all CMSs irrespective of whether a medium involves a cholesterol uptake reaction. Author summaryEssential genes are indispensable genes for cells to grow and proliferate under certain physiological condition. Identifying essential genes in genome-scale metabolic networks of cancer cells is a time-consuming process. We develop an anticancer target discovery platform for identifying essential genes that conduct cell death when the genes of cancer cells are deleted. Meanwhile, the essential genes are also inactive on their healthy cells to maintain their cell viability and smaller metabolic alterations. We use fuzzy set theory to measure metabolic deviation of the perturbation of normal cells relative to healthy and cancer templates towards predicting side effects for treatment of each identified gene. The platform can identify essential genes, metabolites and reactions for treating five consensus molecular subtypes (CMS) of colorectal cancers with using various media. We discovered that most targets affected all five CMSs and that some genes belonged to a CMS-specific model. We found that the genes in the cholesterol biosynthetic pathway are nonessential for the cells that be compensated by a cholesterol uptake reaction from a medium. Furthermore, CRLS1 was revealed as an essential gene for all CMS colorectal cancer in a medium-independent manner that is unrelated to a cholesterol uptake reaction.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Cheng, C.-T., Lai, J.-M., Chang, P. M.-H., Hong, Y.-R., Huang, C.-Y. F., Wang, F.-S.. 2022-10-12. Identifying Essential Genes in Genome-Scale Metabolic Models of Consensus Molecular Subtypes of Colorectal Cancer. https://doi.org/10.1101/2022.10.04.510777

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Limit-pushing overexpression reveals constraints on protein abundance

Proteins are often classified as toxic or non-toxic without measuring the abundance reached, leaving constraints on tolerable protein abundance unresolved. We established a limit-pushing approach in Saccharomyces cerevisiae combining strong inducible expression with gTOW-mediated high-copy selection to counteract copy-number compensation while measuring protein abundance and growth. Nearly all of approximately 80 chromosome I proteins severely inhibited growth or reduced viability at sufficiently high abundance. We established IE50, the expression level associated with a 50% reduction in growth rate, to quantify their widely varying overexpression tolerance. IE50 was positively associated with predicted structural order and cytoplasmic localization propensity and negatively associated with sulphur content. Single-cell imaging linked higher tolerance to proteins remaining cytoplasmic without becoming aggregation-positive and revealed abundance-dependent changes in localization and organelle morphology. At extreme abundance, Fun12, Nup60, and Pex22 generated distinct large-scale intracellular states through specific sequence regions. These findings establish overexpression toxicity as a quantitative property linked to protein characteristics and reveal both constraints on tolerable abundance and sequence-dependent capacities for intracellular organization.

systems biology↗

Accessing Enzyme Kinetic Data and Prediction Methods at Scale

Enzyme kinetic parameters inform metabolic models, yet experimental measurements are sparse. A growing body of work predicts them from protein and substrate features, but software fragmentation hinders adoption, so downstream tools lock into the most accessible method. We present OpenKinetics Predictor (at predictor.openkinetics.org), an open-source platform integrating thirteen methods in isolated environments behind one interface. The platform optionally reports similarity between query proteins and each method's training data to contextualise reliability. A common featurisation-prediction abstraction keeps it extensible, and independent parties, including original authors, contributed many methods. We pair it with a data portal (at data.openkinetics.org) that exposes CatLog, a curated kinetic dataset, with precomputed embeddings, predicted binding sites, and standardised splits. Both offer a web interface and an API, and the GECKO modelling toolbox already calls the predictor API. As a case study, we predict across an E. coli model and find inter-predictor agreement varies with metabolic context and data availability.

systems biology↗

A thermoregulatory design principle for transitions into hypometabolism

Mammals entering torpor or hibernation undergo an abrupt transition from normothermia to hypothermia, yet how thermoregulation enables this switch remains poorly understood. Here, we identify dynamical signatures that precede these transitions and a mathematical principle that can generate them. In fasting-induced torpor in mice, body-temperature fluctuations increased before torpor onset, providing an early-warning signal that tracked proximity to the transition better than temperature decline alone. A heat-balance model showed that reducing how strongly the effective heat-loss coefficient depends on body temperature reorganizes thermoregulatory stability, allowing a low-temperature equilibrium to emerge while the normothermic state remains stable. This organization is consistent with a symmetry-broken pitchfork involving a saddle-node. Similar increases in temperature fluctuations preceded hibernation onset in hamsters. These findings link pre-transition temperature dynamics to changes in the underlying thermoregulatory landscape and provide a framework for detecting and understanding transitions from normothermia to hypothermia.

systems biology↗