Skip to main navigation Skip to search Skip to main content

An iterative strategy for pattern discovery in high-dimensional data sets

  • SUNY Buffalo

Research output: Contribution to conferencePaperpeer-review

11 Scopus citations

Abstract

High-dimensional data representation in which each data item (termed target object) is described by many features, is a necessary component of many applications. For example, in DNA microarrays, each sample (target object) is represented by thousands of genes as features. Pattern discovery of target objects presents interesting but also very challenging problems. The data sets are typically not task-specific, many features are irrelevant or redundant and should be pruned out or filtered for the purpose of classifying target objects to find empirical pattern. Uncertainty about which features are relevant makes it difficult to construct an informative feature space. This paper proposes an iterative strategy for pattern discovery in high-dimensional data sets. In this approach, the iterative process consists of two interactive components: discovering patterns within target objects and pruning irrelevant features. The performance of the proposed method with various real data sets is also illustrated.

Original languageEnglish
Pages10-17
Number of pages8
DOIs
StatePublished - 2002
EventProceedings of the Eleventh International Conference on Information and Knowledge Management (CIKM 2002) - McLean, VA, United States
Duration: Nov 4 2002Nov 9 2002

Conference

ConferenceProceedings of the Eleventh International Conference on Information and Knowledge Management (CIKM 2002)
Country/TerritoryUnited States
CityMcLean, VA
Period11/4/0211/9/02

Keywords

  • Empirical pattern
  • Feature
  • Iterative
  • Microarray
  • Unsupervise

Fingerprint

Dive into the research topics of 'An iterative strategy for pattern discovery in high-dimensional data sets'. Together they form a unique fingerprint.

Cite this