PSI-2: structural genomics to cover protein domain family space.

TitlePSI-2: structural genomics to cover protein domain family space.
Publication TypeJournal Article
Year of Publication2009
AuthorsDessailly, BH, Nair, R, Jaroszewski, L, J Fajardo, E, Kouranov, A, Lee, D, Fiser, A, Godzik, A, Rost, B, Orengo, C
JournalStructure
Volume17
Issue6
Pagination869-81
Date Published2009 Jun 10
ISSN1878-4186
KeywordsAnimals, Computational Biology, Genomics, Humans, Multigene Family, Protein Conformation, Protein Structure, Tertiary, Proteins, Proteomics, Sequence Analysis, Protein
Abstract

One major objective of structural genomics efforts, including the NIH-funded Protein Structure Initiative (PSI), has been to increase the structural coverage of protein sequence space. Here, we present the target selection strategy used during the second phase of PSI (PSI-2). This strategy, jointly devised by the bioinformatics groups associated with the PSI-2 large-scale production centers, targets representatives from large, structurally uncharacterized protein domain families, and from structurally uncharacterized subfamilies in very large and diverse families with incomplete structural coverage. These very large families are extremely diverse both structurally and functionally, and are highly overrepresented in known proteomes. On the basis of several metrics, we then discuss to what extent PSI-2, during its first 3 years, has increased the structural coverage of genomes, and contributed structural and functional novelty. Together, the results presented here suggest that PSI-2 is successfully meeting its objectives and provides useful insights into structural and functional space.

DOI10.1016/j.str.2009.03.015
Alternate JournalStructure
PubMed ID19523904
PubMed Central IDPMC2920419
Grant ListU54 GM074942-040003 / GM / NIGMS NIH HHS / United States