Borderline-SMOTE: a new over-sampling method in imbalanced data sets learning

ADVANCES IN INTELLIGENT COMPUTING, PT 1, PROCEEDINGS(2005)

引用 4509|浏览11
暂无评分
摘要
In recent years, mining with imbalanced data sets receives more and more attentions in both theoretical and practical aspects. This paper introduces the importance of imbalanced data sets and their broad application domains in data mining, and then summarizes the evaluation metrics and the existing methods to evaluate and solve the imbalance problem. Synthetic minority over-sampling technique (SMOTE) is one of the over-sampling methods addressing this problem. Based on SMOTE method, this paper presents two new minority over-sampling methods, borderline-SMOTE1 and borderline-SMOTE2, in which only the minority examples near the borderline are over-sampled. For the minority class, experiments show that our approaches achieve better TP rate and F-value than SMOTE and random over-sampling methods.
更多
查看译文
关键词
synthetic minority,over-sampling method,data mining,minority class,smote method,new over-sampling method,minority example,new minority,imbalanced data set,random over-sampling method,imbalance problem,information extraction,sampling methods,artificial intelligence,random sampling,sampling technique,metric,data analysis
AI 理解论文
溯源树
样例
生成溯源树,研究论文发展脉络
Chat Paper
正在生成论文摘要