MISS: a non-linear methodology based on mutual information for genetic association studies in both population and sib-pairs analysis.

Bioinformatics

Institut de Bioenginyeria de Catalunya, Departament d'Enginyeria de Sistemes, Automàtica i Informàtica Industrial, Universitat Politècnica de Catalunya, Pau Gargallo 5, 08028 Barcelona, Spain.

Published: August 2010

Motivation: Finding association between genetic variants and phenotypes related to disease has become an important vehicle for the study of complex disorders. In this context, multi-loci genetic association might unravel additional information when compared with single loci search. The main goal of this work is to propose a non-linear methodology based on information theory for finding combinatorial association between multi-SNPs and a given phenotype.

Results: The proposed methodology, called MISS (mutual information statistical significance), has been integrated jointly with a feature selection algorithm and has been tested on a synthetic dataset with a controlled phenotype and in the particular case of the F7 gene. The MISS methodology has been contrasted with a multiple linear regression (MLR) method used for genetic association in both, a population-based study and a sib-pairs analysis and with the maximum entropy conditional probability modelling (MECPM) method, which searches for predictive multi-locus interactions. Several sets of SNPs within the F7 gene region have been found to show a significant correlation with the FVII levels in blood. The proposed multi-site approach unveils combinations of SNPs that explain more significant information of the phenotype than their individual polymorphisms. MISS is able to find more correlations between SNPs and the phenotype than MLR and MECPM. Most of the marked SNPs appear in the literature as functional variants with real effect on the protein FVII levels in blood.

Availability: The code is available at http://sisbio.recerca.upc.edu/R/MISS_0.2.tar.gz

Download full-text PDF

Source
http://dx.doi.org/10.1093/bioinformatics/btq273DOI Listing

Publication Analysis

Top Keywords

genetic association
12
non-linear methodology
8
methodology based
8
sib-pairs analysis
8
fvii levels
8
association
5
based mutual
4
genetic
4
mutual genetic
4
association studies
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!