The Community for Technology Leaders
RSS Icon
Subscribe
Issue No.02 - February (1991 vol.13)
pp: 175-184
ABSTRACT
<p>A methodology for clustering data in which a distance metric or similarity function is not used is described. Instead, clusterings are optimized based on their intended function: the accurate prediction of properties of the data. The resulting clustering methodology is applicable, without further ad hoc assumptions or transformations of the data, (1) when features are heterogeneous (both discrete and continuous) and not combinable, (2) where some data points have missing feature values, and (3) where some features are irrelevant, i.e. have large variance but little correlation with other features. Further, it provides an integral measure of the quality of the resulting clustering. A clustering program, RIFFLE, has been implemented in line with this approach, and experiments with synthetic and real data show that the clustering is, in many respects, superior to traditional methods.</p>
INDEX TERMS
statistical analysis; pattern recognition; optimisation; clustering; distance metric; similarity function; RIFFLE; optimisation; pattern recognition; statistical analysis
CITATION
G. Matthews, J. Hearne, "Clustering Without a Metric", IEEE Transactions on Pattern Analysis & Machine Intelligence, vol.13, no. 2, pp. 175-184, February 1991, doi:10.1109/34.67646
6 ms
(Ver 2.0)

Marketing Automation Platform Marketing Automation Tool