bioRxiv Science⌕ Search

Biology subjects

Chan, E. T.

Publications and source records attributed to Chan, E. T..

2 recordsLinked to original sources

The ENCODE Uniform Analysis Pipelines

The Encyclopedia of DNA elements (ENCODE) project is a collaborative effort to create a comprehensive catalog of functional elements in the human genome. The current database comprises more than 19000 functional genomics experiments across more than 1000 cell lines and tissues using a wide array of experimental techniques to study the chromatin structure, regulatory and transcriptional landscape of the Homo sapiens and Mus musculus genomes. All experimental data, metadata, and associated computational analyses created by the ENCODE consortium are submitted to the Data Coordination Center (DCC) for validation, tracking, storage, and distribution to community resources and the scientific community. The ENCODE project has engineered and distributed uniform processing pipelines in order to promote data provenance and reproducibility as well as allow interoperability between genomic resources and other consortia. All data files, reference genome versions, software versions, and parameters used by the pipelines are captured and available via the ENCODE Portal. The pipeline code, developed using Docker and Workflow Description Language (WDL; https://openwdl.org/) is publicly available in GitHub, with images available on Dockerhub (https://hub.docker.com), enabling access to a diverse range of biomedical researchers. ENCODE pipelines maintained and used by the DCC can be installed to run on personal computers, local HPC clusters, or in cloud computing environments via Cromwell. Access to the pipelines and data via the cloud allows small labs the ability to use the data or software without access to institutional compute clusters. Standardization of the computational methodologies for analysis and quality control leads to comparable results from different ENCODE collections - a prerequisite for successful integrative analyses. Database URL: https://www.encodeproject.org/

bioinformatics↗

Deciphering the Molecular Mechanism of HCV Protease Inhibitor Fluorination as a General Approach to Avoid Drug Resistance

Third generation Hepatitis C virus (HCV) NS3/4A protease inhibitors (PIs), glecaprevir and voxilaprevir, are highly effective across genotypes and against many resistant variants. Unlike earlier PIs, these compounds have fluorine substitutions on the P2-P4 macrocycle and P1 moieties. Fluorination has long been used in medicinal chemistry as a strategy to improve physicochemical properties and potency. However, the molecular basis by which fluorination improves potency and resistance profile of HCV NS3/4A PIs is not well understood. To systematically analyze the contribution of fluorine substitutions to inhibitor potency and resistance profile, we used a multi-disciplinary approach involving inhibitor design and synthesis, enzyme inhibition assays, co-crystallography, and structural analysis. A panel of inhibitors in matched pairs were designed with and without P4 cap fluorination, tested against WT protease and the D168A resistant variant, and a total of 22 high-resolution co-crystal structures were determined. While fluorination did not significantly improve potency against the WT protease, PIs with fluorinated P4 caps retained much better potency against the D168A protease variant. Detailed analysis of the co-crystal structures revealed that PIs with fluorinated P4 caps can sample alternate binding conformations that enable adapting to structural changes induced by the D168A substitution. Our results elucidate molecular mechanisms of fluorine-specific inhibitor interactions that can be leveraged in avoiding drug resistance.

biochemistry↗