bioRxiv · 10.1101/578500
pysradb: A Python package to query next-generation sequencing metadata and data from NCBI Sequence Read Archive
Abstract
NCBIs Sequence Read Archive (SRA) is the primary archive of next-generation sequencing datasets. SRA makes metadata and raw sequencing data available to the research community to encourage reproducibility, and to provide avenues for testing novel hypotheses on publicly available data. However, existing methods to programmatically access these data are limited. We introduce a Python package pysradb that provides a collection of command line methods to query and download metadata and data from SRA utilizing the curated metadata database available through the SRAdb project. We demonstrate the utility of pysradb on multiple use cases for searching and downloading SRA datasets. It is available freely at https://github.com/saketkc/pysradb.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Choudhary, S.. 2019-03-16. pysradb: A Python package to query next-generation sequencing metadata and data from NCBI Sequence Read Archive. https://doi.org/10.1101/578500
Cite the original work for its findings. Save a collection to share your selection of sources.