bioRxiv · 10.1101/2023.11.07.565968
Statistical framework for calling allelic imbalance in high-throughput sequencing data
Abstract
High-throughput sequencing facilitates large-scale studies of gene regulation and allows tracing the associations of individual genomic variants with changes in gene expression. Compared to classic association studies, allelic imbalance at heterozygous variants captures the functional effects of the regulatory genome variation with smaller sample sizes and higher sensitivity. Yet, the identification of allele-specific events from allelic read counts remains non-trivial due to multiple sources of technical and biological variability, which induce data-dependent biases and overdispersion. Here we present MIXALIME, a novel computational framework for calling allele-specific events in diverse omics data with a repertoire of statistical models accounting for read mapping bias and copy-number variation. We benchmark MIXALIME against existing tools and demonstrate its practical usage by constructing an atlas of allele-specific chromatin accessibility, UDACHA, from thousands of available datasets obtained from diverse cell types. Availabilityhttps://github.com/autosome-ru/MixALime, https://udacha.autosome.org
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Andrey Buyan, A., Meshcheryakov, G., Safronov, V., Abramov, S., Boytsov, A., Nozdrin, V., Baulin, E. F., Kolmykov, S., Vierstra, J., Kolpakov, F., Makeev, V. J., Kulakovskiy, I. V.. 2023-11-09. Statistical framework for calling allelic imbalance in high-throughput sequencing data. https://doi.org/10.1101/2023.11.07.565968
Cite the original work for its findings. Save a collection to share your selection of sources.