Abstract
Underpinning a significant amount of the mass quantities of data, virtualization technology is a key element of utility cloud and an area in which monitoring is a special challenge. The monitoring of large, complex systems requires high accuracy, low latency, and near-real-time fault detection and anomaly analysis along with optimization enactment and actions for corrections. For this paper, we investigated a fine-grained fault-tolerance mechanism with newly proposed algorithms for the analysis of large datasets that are based on the Hadoop MapReduce platform, and we implement a Naïve Bayes Classifier (NBC) algorithm with Hadoop MapReduce to achieve high-performance and efficient classification for the analysis procedure that occurs in virtualization and utility cloud. Evaluation results show that the accuracy of our proposed method using Hadoop MapReduce approaches 89.80% as the size of the data sets increases. We demonstrate that our model is scalable to large data sets of virtual machine (VM) component utilization metrics with increased accuracy, low latency, and machine learning ability.
| Original language | English |
|---|---|
| Pages (from-to) | 185-202 |
| Number of pages | 18 |
| Journal | Journal of Computers (Taiwan) |
| Volume | 29 |
| Issue number | 4 |
| DOIs | |
| State | Published - Aug 2018 |
| Externally published | Yes |
Keywords
- Fault diagnosis
- Hadoop MapReduce
- Naïve bayes classifier
- Utility cloud
- Virtualization
Fingerprint
Dive into the research topics of 'Improving fault diagnosis performance using hadoop mapreduce for efficient classification and analysis of large data sets'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver