Statistical analysis of domains in interacting protein pairs.

Tom M W Nye Carlo Berzuini Walter R Gilks M Madan Babu Sarah A Teichmann

Bioinformatics

Medical Research Council Biostatistics Unit, Cambridge, UK.

Published: April 2005

Motivation: Several methods have recently been developed to analyse large-scale sets of physical interactions between proteins in terms of physical contacts between the constituent domains, often with a view to predicting new pairwise interactions. Our aim is to combine genomic interaction data, in which domain-domain contacts are not explicitly reported, with the domain-level structure of individual proteins, in order to learn about the structure of interacting protein pairs. Our approach is driven by the need to assess the evidence for physical contacts between domains in a statistically rigorous way.

Results: We develop a statistical approach that assigns p-values to pairs of domain superfamilies, measuring the strength of evidence within a set of protein interactions that domains from these superfamilies form contacts. A set of p-values is calculated for SCOP superfamily pairs, based on a pooled data set of interactions from yeast. These p-values can be used to predict which domains come into contact in an interacting protein pair. This predictive scheme is tested against protein complexes in the Protein Quaternary Structure (PQS) database, and is used to predict domain-domain contacts within 705 interacting protein pairs taken from our pooled data set.

Download full-text PDF	Source
http://dx.doi.org/10.1093/bioinformatics/bti086	DOI Listing

Publication Analysis

Top Keywords

interacting protein

protein pairs

physical contacts

domain-domain contacts

pooled data

data set

protein

domains

pairs

contacts

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!