Comparing the performance of supervised classification methods on a multispecies fishery of post-larval galaxiids

A key challenge in clinical proteomics of cancer is the identification of biomarkers that could allow detection, diagnosis and prognosis of the diseases. Recent advances in mass spectrometry and proteomic instrumentations offer unique chance to rapidly identify these markers. These advances pose considerable challenges, similar to those created by microarray-based investigation, for the discovery of pattern of markers from high-dimensional data, specific to each pathologic state (e.g. normal vs cancer). We propose a three-step strategy to select important markers from high-dimensional mass spectrometry data using surface enhanced laser desorption/ionization (SELDI) technology. The first two steps are the selection of the most discriminating biomarkers with a construction of different classifiers. Finally, we compare and validate their performance and robustness using different supervised classification methods such as Support Vector Machine, Linear Discriminant Analysis, Quadratic Discriminant Analysis, Neural Networks, Classification Trees and Boosting Trees. We show that the proposed method is suitable for analysing high-throughput proteomics data and that the combination of logistic regression and Linear Discriminant Analysis outperform other methods tested.

Download Full-text

Supervised Classification Methods for Fake News Identification

Artificial Intelligence and Soft Computing - Lecture Notes in Computer Science ◽

10.1007/978-3-030-61534-5_40 ◽

2020 ◽

pp. 445-454

Author(s):

Thanh Cong Truong ◽

Quoc Bao Diep ◽

Ivan Zelinka ◽

Roman Senkerik

Keyword(s):

Supervised Classification ◽

Classification Methods ◽

Fake News ◽

Supervised Classification Methods

Download Full-text

A Comparison of Semi-Supervised Classification Approaches for Software Defect Prediction

Journal of Intelligent Systems ◽

10.1515/jisys-2013-0030 ◽

2014 ◽

Vol 23 (1) ◽

pp. 75-82 ◽

Cited By ~ 12

Author(s):

Cagatay Catal

Keyword(s):

Supervised Classification ◽

Defect Prediction ◽

Support Vector ◽

Software Defect Prediction ◽

Classification Methods ◽

Data Set ◽

Software Defect ◽

Data Points ◽

Supervised Classification Methods ◽

Prediction Approach

AbstractPredicting the defect-prone modules when the previous defect labels of modules are limited is a challenging problem encountered in the software industry. Supervised classification approaches cannot build high-performance prediction models with few defect data, leading to the need for new methods, techniques, and tools. One solution is to combine labeled data points with unlabeled data points during learning phase. Semi-supervised classification methods use not only labeled data points but also unlabeled ones to improve the generalization capability. In this study, we evaluated four semi-supervised classification methods for semi-supervised defect prediction. Low-density separation (LDS), support vector machine (SVM), expectation-maximization (EM-SEMI), and class mass normalization (CMN) methods have been investigated on NASA data sets, which are CM1, KC1, KC2, and PC1. Experimental results showed that SVM and LDS algorithms outperform CMN and EM-SEMI algorithms. In addition, LDS algorithm performs much better than SVM when the data set is large. In this study, the LDS-based prediction approach is suggested for software defect prediction when there are limited fault data.

Download Full-text