Abstract
Homorepeat sequences, consecutive runs of identical amino acids, are prevalent in eu-karyotic proteins. It has become necessary to annotate and evaluate this feature in entire proteomes. The definition of what constitutes a homorepeat is not fixed, and different research approaches may require different definitions; therefore, flexible approaches to analyze homorepeats in complete pro-teomes are needed. Here, we present polyX2, a fast, simple but tunable script to scan protein datasets for all possible homorepeats. The user can modify the length of the window to scan, the minimum number of identical residues that must be found in the window, and the types of homorepeats to be found.
Author supplied keywords
Cite
CITATION STYLE
Mier, P., & Andrade-Navarro, M. A. (2022). PolyX2: Fast Detection of Homorepeats in Large Protein Datasets. Genes, 13(5). https://doi.org/10.3390/genes13050758
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.