A Comparative Analysis of Various Cluster Detection Techniques for Data Mining

Stream data mining is a popular research area these days. The concept drift detection and drift handling are the biggest challenges of stream data mining. Several drift detection algorithms have been developed which can accurately detect various drifts but have the problem of false-positive drift detection. The false-positive drift detection leads to the performance degradation of the classifier because of unnecessary training in between analyses. Classifier ensemble has shown its efficiency for drift detection, drift handling, and classification. But the ensemble classifiers could not detect the exact position of drift occurrence, so it has to update itself at some fixed interval, which leads to an unnecessary computational burden on the system. Combining the drift detection algorithm with an ensemble classifier can improve the performance and also solve the problems of false-positive drift detection and unnecessary updating of the ensemble classifier. In this paper, a model is proposed that creates a weighted adaptive ensemble classifier by updating it only when a drift detection signal is given by the used drift detection method. The proposed model is evaluated on text-based stream data for sentiment analysis and opinion mining with multiple drift detection algorithms and with multiple classification algorithms as base classifiers for the ensemble. A comparative analysis has been done, and the results have shown the efficiency of the proposed models.

Download Full-text

Spam Mail Detection Using Data Mining: A Comparative Analysis

Smart Intelligent Computing and Applications - Smart Innovation, Systems and Technologies ◽

10.1007/978-981-13-1921-1_56 ◽

2018 ◽

pp. 571-580 ◽

Cited By ~ 1

Author(s):

Soumyabrata Saha ◽

Suparna DasGupta ◽

Suman Kumar Das

Keyword(s):

Data Mining ◽

Comparative Analysis ◽

Using Data

Download Full-text

Comparative Analysis of Data Mining Models for Crop Yield by Using Rainfall and Soil Attributes

2018 Second International Conference on Inventive Communication and Computational Technologies (ICICCT) ◽

10.1109/icicct.2018.8473074 ◽

2018 ◽

Author(s):

Kunal Teeda ◽

Nandini Vallabhaneni ◽

T. Sridevi

Keyword(s):

Data Mining ◽

Comparative Analysis ◽

Crop Yield ◽

Soil Attributes

Download Full-text

A Survey on Major Classification Algorithms and Comparative Analysis of Few Classification Algorithms on Contact Lenses Data Set Using Data Mining Tool

New Trends in Computational Vision and Bio-inspired Computing ◽

10.1007/978-3-030-41862-5_121 ◽

2020 ◽

pp. 1201-1209

Author(s):

Syed Nawaz Pasha ◽

D. Ramesh ◽

Mohammad Sallauddin

Keyword(s):

Data Mining ◽

Comparative Analysis ◽

Contact Lenses ◽

Classification Algorithms ◽

Data Set ◽

Data Mining Tool ◽

Mining Tool ◽

Using Data

Download Full-text

A comparative Analysis of Multiple Regression in Data Mining

Journal of Computer & Information Technology ◽

10.22147/jucit/080601 ◽

2017 ◽

Vol 08 (06) ◽

pp. 37-40

Author(s):

PRIYANKA VERMA ◽

◽

RAJNI KORI ◽

SHIV KUMAR ◽

◽

...

Keyword(s):

Data Mining ◽

Comparative Analysis ◽

Multiple Regression

Download Full-text

Analysis of Classification and Clustering based Novel Class Detection Techniques for Stream Data Mining

International Journal of Engineering Research and ◽

10.17577/ijertv4is100160 ◽

2015 ◽

Vol V4 (10) ◽

Author(s):

Kamini Tandel ◽

Jignasa N. Patel ◽

Keyword(s):

Data Mining ◽

Stream Data ◽

Detection Techniques ◽

Stream Data Mining ◽

Classification And Clustering

Download Full-text

A Comparative Analysis of Data Mining Techniques on Breast Cancer Diagnosis Data using WEKA Toolbox

International Journal of Advanced Computer Science and Applications ◽

10.14569/ijacsa.2020.0110829 ◽

2020 ◽

Vol 11 (8) ◽

Cited By ~ 1

Author(s):

Majdah Alshammari ◽

Mohammad Mezher

Keyword(s):

Breast Cancer ◽

Data Mining ◽

Comparative Analysis ◽

Cancer Diagnosis ◽

Breast Cancer Diagnosis ◽

Data Mining Techniques

Download Full-text

Comparative Analysis of Edge Detection Techniques for Medical Images of Different Body Parts

Data Science and Analytics - Communications in Computer and Information Science ◽

10.1007/978-981-10-8527-7_15 ◽

2018 ◽

pp. 164-176

Author(s):

Bhawna Dhruv ◽

Neetu Mittal ◽

Megha Modi

Keyword(s):

Comparative Analysis ◽

Edge Detection ◽

Medical Images ◽

Body Parts ◽

Detection Techniques

Download Full-text

Distance Based Pattern Driven Mining for Outlier Detection in High Dimensional Big Dataset

ACM Transactions on Management Information Systems ◽

10.1145/3469891 ◽

2022 ◽

Vol 13 (1) ◽

pp. 1-17

Author(s):

Ankit Kumar ◽

Abhishek Kumar ◽

Ali Kashif Bashir ◽

Mamoon Rashid ◽

V. D. Ambeth Kumar ◽

...

Keyword(s):

Data Mining ◽

Comparative Analysis ◽

Outlier Detection ◽

Credit Card ◽

High Dimensional ◽

Work Efficiency ◽

Average Value ◽

Novel Method ◽

Detection Of Outliers ◽

Better Than

Detection of outliers or anomalies is one of the vital issues in pattern-driven data mining. Outlier detection detects the inconsistent behavior of individual objects. It is an important sector in the data mining field with several different applications such as detecting credit card fraud, hacking discovery and discovering criminal activities. It is necessary to develop tools used to uncover the critical information established in the extensive data. This paper investigated a novel method for detecting cluster outliers in a multidimensional dataset, capable of identifying the clusters and outliers for datasets containing noise. The proposed method can detect the groups and outliers left by the clustering process, like instant irregular sets of clusters (C) and outliers (O), to boost the results. The results obtained after applying the algorithm to the dataset improved in terms of several parameters. For the comparative analysis, the accurate average value and the recall value parameters are computed. The accurate average value is 74.05% of the existing COID algorithm, and our proposed algorithm has 77.21%. The average recall value is 81.19% and 89.51% of the existing and proposed algorithm, which shows that the proposed work efficiency is better than the existing COID algorithm.

Download Full-text

Visual Data Mining: A Comparative Analysis of Selected Datasets

Advances in Intelligent Systems and Computing - Intelligent Systems Design and Applications ◽

10.1007/978-3-030-71187-0_35 ◽

2021 ◽

pp. 377-391

Author(s):

Ujunwa Mgboh ◽

Blessing Ogbuokiri ◽

George Obaido ◽

Kehinde Aruleba

Keyword(s):

Data Mining ◽

Comparative Analysis ◽

Visual Data ◽

Visual Data Mining

Download Full-text