Difference between revisions of "Clustering"

Revision as of 20:29, 22 April 2019

Similarity Measures for Clusters:

Compare the numbers of identical and unique item pairs appearing in cluster sets
Achieved by counting the number of item pairs found in both clustering sets (a) as well as the pairs appearing only in the first (b) or the second (c) set.
With this a similarity coefficient, such as the Jaccard index, can be computed. The latter is defined as the size of the intersect divided by the size of the union of two sample sets: a/(a+b+c).
In case of partitioning results, the Jaccard Index measures how frequently pairs of items are joined together in two clustering data sets and how often pairs are observed only in one set.
Related coefficient are the Rand Index and the Adjusted Rand Index. These indices also consider the number of pairs (d) that are not joined together in any of the clusters in both sets

@@ Line 21: / Line 21: @@
 ***[[Hierarchical Clustering;  Agglomerative (HAC) & Divisive (HDC)]]
 ***[[Hierarchical Temporal Memory (HTM)]]
+Similarity Measures for Clusters:
+* Compare the numbers of identical and unique item pairs appearing in cluster sets
+* Achieved by counting the number of item pairs found in both clustering sets (a) as well as the pairs appearing only in the first (b) or the second (c) set.
+* With this a similarity coefficient, such as the Jaccard index, can be computed. The latter is defined as the size of the intersect divided by the size of the union of two sample sets: a/(a+b+c).
+* In case of partitioning results, the Jaccard Index measures how frequently pairs of items are joined together in two clustering data sets and how often pairs are observed only in one set.
+* Related coefficient are the Rand Index and the Adjusted Rand Index. These indices also consider the number of pairs (d) that are not joined together in any of the clusters in both sets
+[http://girke.bioinformatics.ucr.edu/GEN242/mydoc_Rclustering_3.html#example-2 Clustering Algorithms | Data Analysis in Genome Biology]
 <youtube>CtKeHnfK5uA</youtube>
@@ Line 28: / Line 36: @@
 <youtube>ZueoXMgCd1c</youtube>
 <youtube>nk9K2AiFmjE</youtube>
-=== Hierarchical Clustering Analysis (HCA) ===
-<youtube>EQZaSuK-PHs</youtube>
-<youtube>JcfIeaGzF8A</youtube>
-<youtube>7BPLNOMNIXM</youtube>
-<youtube>EUQY3hL38cw</youtube>