Numerical classification of coding sequences

5Citations
Citations of this article
11Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

DNA sequences coding for protein may be represented by counts of nucleotides or codons. A complete reading frame may be abbreviated by its base count, e.g. A76C158G121T74 or with the corresponding codon table, e.g. (AAA)o(AAC)1(AAG)9 ...(TTT)o. We propose that these numerical designations be used to augment current methods of sequence annotation. Because base counts and codon tables do not require revision as knowledge of function evolves, they are well-suited to act as cross-references, for example to identify redundant GenBank entries. These descriptors may be compared, in place of DNA sequences, to extract homologous genes from large databases. This approach permits rapid searching with good selectivity. © 1992 IRL Press at Oxford University Press.

Cite

CITATION STYLE

APA

Collins, D. W., Liu, C. chang, & Jukes, T. H. (1992). Numerical classification of coding sequences. Nucleic Acids Research, 20(6), 1405–1410. https://doi.org/10.1093/nar/20.6.1405

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free