Spectral Clustering and Visualization: A novel Clustering of Fisher's Iris Data Set

David Benson-Putnins,Meagan E Magnoni,Margaret Bonfardin,Daniel Martin

doi:10.1137/10s010752

Abstract

Clustering is the act of partitioning a set of elements into subsets, or clusters, so that elements in the same cluster are, in some sense, similar. Determining an appropriate number of clusters in a particular data set is an important issue in data mining and cluster analysis. Another important issue is visualizing the strength, or connectivity, of clusters. We begin by creating a consensus matrix using multiple runs of the clustering algorithm k-means. This consensus matrix can be interpreted as a graph, which we cluster using two spectral clustering methods: the Fiedler Method and the MinMaxCut Method. To determine if increasing the number of clusters from k to k + 1 is appropriate, we check whether an existing cluster can be split. Finally, we visualize the strength of clusters by using the consensus matrix and the clustering obtained through one of the aforementioned spectral clustering techniques. Using these methods, we then investigate Fisher’s Iris data set. Our methods support the existence of four clusters, instead of the generally accepted three clusters in this data.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Spectral Clustering and Visualization: A novel Clustering of Fisher's Iris Data Set

Abstract

Talk to us

Similar Papers

More From: SIAM Undergraduate Research Online

Lead the way for us

Journal: SIAM Undergraduate Research Online	Publication Date: Jan 1, 2011
Citations: 19

Similar Papers

Implementing Clustering with Weka and R

-

01 Apr 2019
01 Apr 2019

Application of evolution strategy in cluster analysis
Yan Ling ... Jiang Jing-Ping
-
Yan Ling, et. al. Yan Ling ... Jiang Jing-Ping
15 Jun 2004
15 Jun 2004

A Moment-Based Classification Algorithm and Its Comparison to the K-Nearest Neighbours for 2-Dimensional Data
Krzysztof Pawelec ... Artem Lenskiy
-
Krzysztof Pawelec, et. al.Krzysztof Pawelec ... Artem Lenskiy
01 Aug 2016
01 Aug 2016

Elucidating the solution structure of the K-means cost function using energy landscape theory.
L Dicks ... D J Wales
The Journal of Chemical Physics | VOL. 156
L Dicks, et. al.L Dicks ... D J Wales
02 Feb 2022
The Journal of Chemical Physics | VOL. 156

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Spectral Clustering and Visualization: A novel Clustering of Fisher's Iris Data Set

Abstract

Talk to us

Similar Papers

More From: SIAM Undergraduate Research Online