bioRxiv · 10.1101/2024.11.11.623070
Poplar: A Phylogenetics Pipeline
Abstract
MotivationGenerating phylogenetic trees from genomic data is essential in understanding biological systems. Each step of this complex process has received extensive attention in the literature, and has been significantly streamlined over the years. Given the volume of publicly available genetic data, obtaining genomes for a wide selection of known species is straightforward. However, analyzing that same data in order to generate a phylogenetic tree is a multi-step process with legitimate scientific and technical challenges, and often requires a significant input from a domain-area scientist. ResultsWe present Poplar, a new, streamlined computational pipeline, to address the computational logistical issues that arise when constructing phylogenetic trees. It provides a framework that runs state-of-the-art software for essential steps in the phylogenetic pipeline, beginning from a genome with or without an annotation, and resulting in a species tree. Running Poplar requires no external databases. In the execution, it enables parallelism for execution for clusters and cloud computing. The trees generated by Poplar match closely with state-of-the-art published trees. The usage and performance of Poplar is far simpler and quicker than manually running a phylogenetic pipeline. Availability and ImplementationFreely available on GitHub at https://github.com/sandialabs/poplar. Implemented using Python and supported on Linux. Supplementary InformationNewick versions of the reference and generated trees.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Krishnakumar, R., Koning, E.. 2024-11-14. Poplar: A Phylogenetics Pipeline. https://doi.org/10.1101/2024.11.11.623070
Cite the original work for its findings. Save a collection to share your selection of sources.