☆ 4.7 Article

Knowledge acquisition through information granulation for imbalanced data

EXPERT SYSTEMS WITH APPLICATIONS (2006)

Journal

EXPERT SYSTEMS WITH APPLICATIONS

Volume 31, Issue 3, Pages 531-541

Publisher

PERGAMON-ELSEVIER SCIENCE LTD

DOI: 10.1016/j.eswa.2005.09.082

Keywords

information granulation; fuzzy ART; granular computing; knowledge acquisition; imbalanced data

Ask authors/readers for more resources

Protocol

Community support

Reagent

Community support

Abstract

When learning from imbalanced/skewed data, which almost all the instances are labeled as one class while far few instances are labeled as the other class, traditional machine learning algorithms tend to produce high accuracy over the majority class but poor predictive accuracy over the minority class. This paper proposes a novel method called 'knowledge acquisition via information granulation' (KAIG) model which not only can remove some unnecessary details and provide a better insight into the essence of data but also effectively solve 'class imbalance' problems. In this model, the homogeneity index (H-index) and the undistinguishable ratio (U-ratio) are successfully introduced to determine a suitable level of granularity. We also developed the concept of sub-attributes to describe granules and tackle the overlapping among granules. Seven data sets from UCI data bank, including one imbalanced diagnosis data (pima-Indians-diabetes), are provided to evaluate the effectiveness of KAIG model. By using different performance indexes, overall accuracy, G-mean and Receiver Operation Characteristic (ROC) curve, the experimental results comparing with C4.5 and Support Vector Machine (SVM) demonstrate the superiority of our method. (c) 2005 Elsevier Ltd. All rights reserved.

Knowledge acquisition through information granulation for imbalanced data

Journal

EXPERT SYSTEMS WITH APPLICATIONS

Publisher

PERGAMON-ELSEVIER SCIENCE LTD

Keywords

Categories

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

Knowledge acquisition through information granulation for imbalanced data

Journal

EXPERT SYSTEMS WITH APPLICATIONS

Publisher

PERGAMON-ELSEVIER SCIENCE LTD

Keywords

Categories

Ask authors/readers for more resources

Protocol

Reagent

Authors

I am an author on this paper

Reviews

Primary Rating

Secondary Ratings

Novelty

Significance

Scientific rigor

Rate this paper

Recommended

Export Citation

Share Paper