Hierarchical clustering is a simple and reproducible technique to rearrange data of multiple variables and sample units and visualize possible groups in the data. Despite the name, hierarchical clustering does not provide clusters automatically, and "tree-cutting" procedures are often used to identify subgroups in the data by cutting the dendrogram that represents the similarities among groups used in the agglomerative procedure. We introduce a resampling-based technique that can be used to identify cut-points of a dendrogram with a significance level based on a reference distribution for the heights of the branch points. The evaluation on synthetic data shows that the technique is robust in a variety of situations. An example with real biomarker data from the Long Life Family Study shows the usefulness of the method.

Download full-text PDF

Source
http://www.ncbi.nlm.nih.gov/pmc/articles/PMC4976109PMC
http://dx.doi.org/10.3389/fgene.2016.00144DOI Listing

Publication Analysis

Top Keywords

hierarchical clustering
12
data
5
detection groups
4
groups hierarchical
4
clustering resampling
4
resampling hierarchical
4
clustering simple
4
simple reproducible
4
reproducible technique
4
technique rearrange
4

Similar Publications

Want AI Summaries of new PubMed Abstracts delivered to your In-box?

Enter search terms and have AI summaries delivered each week - change queries or unsubscribe any time!