bioRxiv · 10.1101/2020.04.18.048439
ProtTox: Toxin identification from Protein Sequences
Abstract
Toxin classification of protein sequences is a challenging task with real world applications in healthcare and synthetic biology. Due to an ever expanding database of proteins and the inordinate cost of manual annotation, automated machine learning based approaches are crucial. Approaches need to overcome challenges of homology, multi-functionality, and structural diversity among proteins in this task. We propose a novel deep learning based method ProtTox, that aims to address some of the shortcomings of previous approaches in classifying proteins as toxins or not. Our method achieves a performance of 0.812 F1-score which is about 5% higher than the closest performing baseline.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Datta, D., Muthiah, S., Butler, P., Islam, M. R., Warren, A., Ramakrishnan, N.. 2020-04-20. ProtTox: Toxin identification from Protein Sequences. https://doi.org/10.1101/2020.04.18.048439
Cite the original work for its findings. Save a collection to share your selection of sources.