Discovering microproteins: making the most of ribosome profiling data

Sonia Chothani; Lena Ho; Sebastian Schafer; Owen Rackham

ArticleOPEN ACCESS

Discovering microproteins: making the most of ribosome profiling data

RNA Biology

DOI: 10.1080/15476286.2023.2279845

16Citations

24Readers

Abstract

Building a reference set of protein-coding open reading frames (ORFs) has revolutionized biological process discovery and understanding. Traditionally, gene models have been confirmed using cDNA sequencing and encoded translated regions inferred using sequence-based detection of start and stop combinations longer than 100 amino-acids to prevent false positives. This has led to small ORFs (smORFs) and their encoded proteins left un-annotated. Ribo-seq allows deciphering translated regions from untranslated irrespective of the length. In this review, we describe the power of Ribo-seq data in detection of smORFs while discussing the major challenge posed by data-quality, -depth and -sparseness in identifying the start and end of smORF translation. In particular, we outline smORF cataloguing efforts in humans and the large differences that have arisen due to variation in data, methods and assumptions. Although current versions of smORF reference sets can already be used as a powerful tool for hypothesis generation, we recommend that future editions should consider these data limitations and adopt unified processing for the community to establish a canonical catalogue of translated smORFs.

Author supplied keywords

Cite

CITATION STYLE

APA

Chothani, S., Ho, L., Schafer, S., & Rackham, O. (2023). Discovering microproteins: making the most of ribosome profiling data. RNA Biology. Taylor and Francis Ltd. https://doi.org/10.1080/15476286.2023.2279845

Discovering microproteins: making the most of ribosome profiling data

Abstract

Author supplied keywords

Cite

Register to see more suggestions