Knowledge discovery in software defect datasets using learning algorithms
DOI:
https://doi.org/10.37591/josettt.v5i2.1748Abstract
In this paper, the learning impact on various classification models were studied which were built using binary class-imbalanced data. Before the learning process, some preprocessing techniques were applied to training datasets for removing the redundancy. Nowadays feature selection and sampling techniques become an essential tool for many data mining task because learning algorithms do not perform well with defective datasets, dimensionality reduction problem arises. Sampling technique is also used to reduce the harmful effects of imbalanced data on prediction models. Two experiments are introduced in this paper are: (1) training on original data with selected features, includes AdaBoost and SVM (support vector machines) as classifiers (2) training on balanced data with chosen features, includes AdaBoost and random forest as the classifier. The classification models were compared over two different schemas. The results demonstrate that the classification models over the selected feature in the balanced format are outperforming the classification model built without balancing (classification over imbalanced data).
Keywords: SDP, SMOTE, AdaBoost, SVM, RF
Cite this Article
Meetesh Nevendra, Pradeep Singh. Knowledge Discovery in Software Defect Datasets Using Learning Algorithms. Journal of Software Engineering Tools & Technology Trends. 2018; 5(2): 18–26p.
Downloads
Published
Issue
Section
License
Declaration and Copyright Transfer Form
(to be completed by authors)
I/ We, the undersigned author(s) of the submitted manuscript, hereby declare, that the above manuscript which is submitted for publication in the STM Journals(s), is not published already in part or whole (except in the form of abstract) in any journal or magazine for private or public circulation, and, is not under consideration of publication elsewhere.
- I/We will not withdraw the manuscript after 1 week of submission as I have read the Author Guidelines and will adhere to the guidelines.
- I/We Author(s ) have niether given nor will give this manuscript elsewhere for publishing after submitting in STM Journal(s).
- I/ We have read the original version of the manuscript and am/ are responsible for the thought contents embodied in it. The work dealt in the manuscript is my/ our own, and my/ our individual contribution to this work is significant enough to qualify for authorship.
- I/We also agree to the authorship of the article in the following order:
Author’s name
1. ________________
2. ________________
3. ________________
4. ________________
| We Author(s) tick this box and would request you to consider it as our signature as we agree to the terms of this Copyright Notice, which will apply to this submission if and when it is published by this journal. |