Reference list of 265 sources used for the discovery of relationships between data clusters and metadata properties

Visual cluster analysis provides valuable tools that help analysts to understand large data sets in terms of representative clusters and relationships thereof. Often, the found clusters are to be understood in context of belonging categorical, numerical or textual metadata which are given for the da...

Full description

Bibliographic Details
Main Authors:	Bernard, Jürgen, Ruppert, Tobias, Scherer, Maximilian, Schreck, Tobias, Kohlhammer, Jörn
Format:	Dataset
Language:	English
Published:	PANGAEA 2012
Subjects:	Alaska USA Antarctica Australia AWIPEV AWIPEV_based BAR Barrow BER Bermuda BOU Boulder Brasilia Brasilia City Distrito Federal Brazil BRB CAB Cabauw Canada CAR Carpentras Chesapeake Light CLH Colorado United States of America Cosmonauts Sea DAR Darwin Dronning Maud Land E13 France Georg von Neumayer Germany GVN Israel Japan KWA Kwajalein LIN Lindenberg MAN Momote Monitoring station MONS NAU Nauru Nauru Island Neumayer Antarc* Cosmonauts sea
Online Access:	https://doi.pangaea.de/10.1594/PANGAEA.785666 https://doi.org/10.1594/PANGAEA.785666

Description
Summary:	Visual cluster analysis provides valuable tools that help analysts to understand large data sets in terms of representative clusters and relationships thereof. Often, the found clusters are to be understood in context of belonging categorical, numerical or textual metadata which are given for the data elements. While often not part of the clustering process, such metadata play an important role and need to be considered during the interactive cluster exploration process. Traditionally, linked-views allow to relate (or loosely speaking: correlate) clusters with metadata or other properties of the underlying cluster data. Manually inspecting the distribution of metadata for each cluster in a linked-view approach is tedious, specially for large data sets, where a large search problem arises. Fully interactive search for potentially useful or interesting cluster to metadata relationships may constitute a cumbersome and long process. To remedy this problem, we propose a novel approach for guiding users in discovering interesting relationships between clusters and associated metadata. Its goal is to guide the analyst through the potentially huge search space. We focus in our work on metadata of categorical type, which can be summarized for a cluster in form of a histogram. We start from a given visual cluster representation, and compute certain measures of interestingness defined on the distribution of metadata categories for the clusters. These measures are used to automatically score and rank the clusters for potential interestingness regarding the distribution of categorical metadata. Identified interesting relationships are highlighted in the visual cluster representation for easy inspection by the user. We present a system implementing an encompassing, yet extensible, set of interestingness scores for categorical metadata, which can also be extended to numerical metadata. Appropriate visual representations are provided for showing the visual correlations, as well as the calculated ranking scores. Focusing on ...

Reference list of 265 sources used for the discovery of relationships between data clusters and metadata properties

Similar Items