Stefan Budach scite author profile

SummaryConvolutional neural networks (CNNs) have been shown to perform exceptionally well in a variety of tasks, including biological sequence classification. Available implementations, however, are usually optimized for a particular task and difficult to reuse. To enable researchers to utilize these networks more easily, we implemented pysster, a Python package for training CNNs on biological sequence data. Sequences are classified by learning sequence and structure motifs and the package offers an automated hyper-parameter optimization procedure and options to visualize learned motifs along with information about their positional and class enrichment. The package runs seamlessly on CPU and GPU and provides a simple interface to train and evaluate a network with a handful lines of code. Using an RNA A-to-I editing dataset and cross-linking immunoprecipitation (CLIP)-seq binding site sequences, we demonstrate that pysster classifies sequences with higher accuracy than previous methods, such as GraphProt or ssHMM, and is able to recover known sequence and structure motifs.Availability and implementationpysster is freely available at https://github.com/budach/pysster.Supplementary information Supplementary data are available at Bioinformatics online.

show abstract

pysster: Classification of Biological Sequences by Learning Sequence and Structure Motifs with Convolutional Neural Networks

Budach

Marsico

2017

Preprint

View full text Add to dashboard Cite

Summary: Convolutional neural networks (CNNs) have been shown to perform exceptionally well in a variety of tasks, including biological sequence classification. Available implementations, however, are usually optimized for a particular task and difficult to reuse. To enable researchers to utilize these networks more easily we implemented pysster, a Python package for training CNNs on biological sequence data. Sequences are classified by learning sequence and structure motifs and the package offers an automated hyper-parameter optimization procedure and options to visualize learned motifs along with information about their positional and class enrichment. The package runs seamlessly on CPU and GPU and provides a simple interface to train and evaluate a network with a handful lines of code. Using an RNA A-to-I editing data set and CLIP-seq binding site sequences we demonstrate that pysster classifies sequences with higher accuracy than other methods and is able to recover known sequence and structure motifs. Availability: pysster is freely available at https://github.com/budach/pysster. Contact:

show abstract

Generic accelerated sequence alignment in SeqAn using vectorization and multi-threading

Rahn

Budach

Costanza

et al. 2018

View full text Add to dashboard Cite

show abstract

Graph Convolutional Networks Improve the Prediction of Cancer Driver Genes

Schulte-Sasse

Budach

Hnisz

et al. 2019

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Stefan Budach

Integration of multiomics data with graph convolutional networks to identify new cancer genes and their associated molecular mechanisms

pysster: classification of biological sequences by learning sequence and structure motifs with convolutional neural networks

pysster: Classification of Biological Sequences by Learning Sequence and Structure Motifs with Convolutional Neural Networks

Generic accelerated sequence alignment in SeqAn using vectorization and multi-threading

Graph Convolutional Networks Improve the Prediction of Cancer Driver Genes

Contact Info

Product

Resources

About