Nevill-Manning CG, et al. (1998)

Reference: Nevill-Manning CG, et al. (1998) Highly specific protein sequence motifs for genome analysis. Proc Natl Acad Sci U S A 95(11):5865-71

Reference Help

Abstract

We present a method for discovering conserved sequence motifs from families of aligned protein sequences. The method has been implemented as a computer program called EMOTIF (http://motif. stanford.edu/emotif). Given an aligned set of protein sequences, EMOTIF generates a set of motifs with a wide range of specificities and sensitivities. EMOTIF also can generate motifs that describe possible subfamilies of a protein superfamily. A disjunction of such motifs often can represent the entire superfamily with high specificity and sensitivity. We have used EMOTIF to generate sets of motifs from all 7,000 protein alignments in the BLOCKS and PRINTS databases. The resulting database, called IDENTIFY (http://motif. stanford.edu/identify), contains more than 50,000 motifs. For each alignment, the database contains several motifs having a probability of matching a false positive that range from 10(-10) to 10(-5). Highly specific motifs are well suited for searching entire proteomes, while generating very few false predictions. IDENTIFY assigns biological functions to 25-30% of all proteins encoded by the Saccharomyces cerevisiae genome and by several bacterial genomes. In particular, IDENTIFY assigned functions to 172 of proteins of unknown function in the yeast genome.

Download Citation (.nbib)

Reference Type: Journal Article | Research Support, Non-U.S. Gov't | Research Support, U.S. Gov't, P.H.S.
Authors: Nevill-Manning CG, Wu TD, Brutlag DL
Primary Lit For
Additional Lit For
Review For

Gene Ontology Annotations

Evidence ID	Analyze ID	Gene/Complex	Systematic Name/Complex Accession	Qualifier	Gene Ontology Term ID	Gene Ontology Term	Aspect	Annotation Extension	Evidence	Method	Source	Assigned On	Reference

Download (.txt)

Analyze

Phenotype Annotations

Evidence ID	Analyze ID	Gene	Gene Systematic Name	Phenotype	Experiment Type	Experiment Type Category	Mutant Information	Strain Background	Chemical	Details	Reference

Download (.txt)

Analyze

Disease Annotations

Evidence ID	Analyze ID	Gene	Gene Systematic Name	Disease Ontology Term	Disease Ontology Term ID	Qualifier	Evidence	Method	Source	Assigned On		Reference

Download (.txt)

Analyze

Regulation Annotations

Evidence ID	Analyze ID	Regulator	Regulator Systematic Name	Target	Target Systematic Name	Direction	Regulation of	Happens During	Regulator Type	Direction	Regulation Of	Happens During	Method	Evidence	Strain Background	Reference

Download (.txt)

Analyze

Post-translational Modifications

				Site		Modification	Modifier	Source	Reference

Download (.txt)

Analyze

Interaction Annotations

Genetic Interactions

Evidence ID	Analyze ID		Interactor	Interactor Systematic Name	Interactor	Interactor Systematic Name	Allele	Assay	Annotation	Action	Phenotype	SGA score	P-value	Source	Reference	Note

Download (.txt)

Analyze

Physical Interactions

Evidence ID	Analyze ID		Interactor	Interactor Systematic Name	Interactor	Interactor Systematic Name	Assay	Annotation	Action	Modification	Source	Reference	Note

Download (.txt)

Analyze

Functional Complementation Annotations

Complement ID	Locus ID	Gene	Species	Gene ID	Strain background	Direction	Details	Source	Reference

Download (.txt)

Analyze

Published Datasets

Evidence ID	Analyze ID		Dataset	Description	Keywords	Number of Conditions	Reference

Download (.txt)

Analyze