Functional insights from the distribution and role of homopeptide repeat-containing proteins

Noel G. Faux, Stephen P. Bottomley, Arthur Lesk, James A. Irving, John R. Morrison, Maria Garcia De La Banda, James C. Whisstock

Research output: Contribution to journalArticle

141 Scopus citations

Abstract

Expansion of "low complex" repeats of amino acids such as glutamine (Poly-Q) is associated with protein misfolding and the development of degenerative diseases such as Huntington's disease. The mechanism by which such regions promote misfolding remains controversial, the function of many repeat-containing proteins (RCPs) remains obscure, and the role (if any) of repeat regions remains to be determined. Here, a Web-accessible database of RCPs is presented. The distribution and evolution of RCPs that contain homopeptide repeats tracts are considered, and the existence of functional patterns investigated. Generally, it is found that while polyamino acid repeats are extremely rare in prokaryotes, several eukaryote putative homologs of prokaryote RCP - involved in important housekeeping processes - retain the repetitive region, suggesting an ancient origin for certain repeats. Within eukarya, the most common uninterrupted amino acid repeats are glutamine, asparagines, and alanine. Interestingly, while poly-Q repeats are found in vertebrates and nonvertebrates, poly-N repeats are only common in more primitive nonvertebrate organisms, such as insects and nematodes. We have assigned function to eukaryote RCPs using Online Mendelian Inheritance in Man (OMIM), the Human Reference Protein Database (HRPD), FlyBase, and Wormpep. Prokaryote RCPs were annotated using BLASTp searches and Gene Ontology. These data reveal that the majority of RCPs are involved in processes that require the assembly of large, multiprotein complexes, such as transcription and signaling.

Original languageEnglish (US)
Pages (from-to)537-551
Number of pages15
JournalGenome Research
Volume15
Issue number4
DOIs
Publication statusPublished - Apr 1 2005

    Fingerprint

All Science Journal Classification (ASJC) codes

  • Genetics
  • Genetics(clinical)

Cite this

Faux, N. G., Bottomley, S. P., Lesk, A., Irving, J. A., Morrison, J. R., De La Banda, M. G., & Whisstock, J. C. (2005). Functional insights from the distribution and role of homopeptide repeat-containing proteins. Genome Research, 15(4), 537-551. https://doi.org/10.1101/gr.3096505