KOBV Portal

Hits per page

hit 1 - 1 | 1 hit

Sorting

Online Resource

Exploratory Analysis of Multiple Omics Datasets Using the Adjusted RV Coefficient

Mayer, Claus-Dieter ; Lorent, Julie ; Horgan, Graham W

Walter de Gruyter GmbH ; 2011

In: Statistical Applications in Genetics and Molecular Biology Vol. 10, No. 1 ( 2011-01-2)

add to watchlist on the watchlist

Details

In: Statistical Applications in Genetics and Molecular Biology, Walter de Gruyter GmbH, Vol. 10, No. 1 ( 2011-01-2)

Abstract: The integration of multiple high-dimensional data sets (omics data) has been a very active but challenging area of bioinformatics research in recent years. Various adaptations of non-standard multivariate statistical tools have been suggested that allow to analyze and visualize such data sets simultaneously. However, these methods typically can deal with two data sets only, whereas systems biology experiments often generate larger numbers of high-dimensional data sets. For this reason, we suggest an explorative analysis of similarity between data sets as an initial analysis steps. This analysis is based on the RV coefficient, a matrix correlation, that can be interpreted as a generalization of the squared correlation from two single variables to two sets of variables. It has been shown before however that the high-dimensionality of the data introduces substantial bias to the RV.We therefore introduce an alternative version, the adjusted RV, which is unbiased in the case of independent data sets. We can also show that in many situations, particularly for very high-dimensional data sets, the adjusted RV is a better estimator than previously RV versions in terms of the mean square error and the power of the independence test based on it. We demonstrate the usefulness of the adjusted RV by applying it to data set of 19 different multivariate data sets from a systems biology experiment. The pairwise RV values between the data sets define a similarity matrix that we can use as an input to a hierarchical clustering or a multi-dimensional scaling. We show that this reveals biological meaningful subgroups of data sets in our study.

Type of Medium: Online Resource

ISSN: 1544-6115 , 2194-6302

URL: Article

DOI: 10.2202/1544-6115.1540

Language: Unknown

Publisher: Walter de Gruyter GmbH

Publication Date: 2011

detail.hit.zdb_id: 2115012-6

Bookmarklink