Search CORE

4 research outputs found

CUSBoost: Cluster-based Under-sampling with Boosting for Imbalanced Classification

Author: Ahmed Sajid
Farid Dewan Md.
Jani Md. Rafsan
Mahbub Asif
Rayhan Farshid
Shatabda Swakkhar
Publication venue: 'Institute of Electrical and Electronics Engineers (IEEE)'
Publication date: 12/12/2017
Field of study

Class imbalance classification is a challenging research problem in data mining and machine learning, as most of the real-life datasets are often imbalanced in nature. Existing learning algorithms maximise the classification accuracy by correctly classifying the majority class, but misclassify the minority class. However, the minority class instances are representing the concept with greater interest than the majority class instances in real-life applications. Recently, several techniques based on sampling methods (under-sampling of the majority class and over-sampling the minority class), cost-sensitive learning methods, and ensemble learning have been used in the literature for classifying imbalanced datasets. In this paper, we introduce a new clustering-based under-sampling approach with boosting (AdaBoost) algorithm, called CUSBoost, for effective imbalanced classification. The proposed algorithm provides an alternative to RUSBoost (random under-sampling with AdaBoost) and SMOTEBoost (synthetic minority over-sampling with AdaBoost) algorithms. We evaluated the performance of CUSBoost algorithm with the state-of-the-art methods based on ensemble learning like AdaBoost, RUSBoost, SMOTEBoost on 13 imbalance binary and multi-class datasets with various imbalance ratios. The experimental results show that the CUSBoost is a promising and effective approach for dealing with highly imbalanced datasets.Comment: CSITSS-201

arXiv.org e-Print Archive

Crossref

iRecSpot-EF: Effective sequence based features for recombination hotspot prediction

Author: Abeysinghe
Baudat
Chen
Chen
Chou
Chowdhury
Cox
Dewan Md Farid
Freund
Friedman
Grigoriev
Hastie
Hey
Izenman
Jeffreys
Jiang
Kabir
Kohavi
Larose
Li
Li
Liaw
Liu
Liu
Liu
Liu
Liu
Liu
Liu
Madigan
Mancera
Md Rafsan Jani
Md Toha Khan Mozlish
Niger Sultana Tahniat
Qiu
Quinlan
Rish
Sajid Ahmed
Swakkhar Shatabda
Vapnik
Zhang
Zhang
Zhang
Zhou
Publication venue: 'Elsevier BV'
Publication date
Field of study

Crossref