K. Anusha scite author profile

With the tremendous increase in the amount of data being generated from variety of sources there is a need of efficient data storage and processing techniques. Some of the sources generating this large amount of data are Weather Sensors, Scientific experiments, etc. This huge voluminous data is termed as BigData. Due to ever-increasing amount of data there is a demand for faster data ingestion and processing. Apache Spark, a dominant processing tool is a publicly available platform for processing outsized data and is mostly intended for iterative machine learning jobs. In this study, an integrated approach i.e., Spark MLlib Clustering on batch weather data stored in Cassandra database is proposed. This helps to analyze our data into number of Clusters which is required and useful for further examination of data. The main idea of this study is to evaluate Batch Processing performance of an integrated approach with two popular clustering algorithms.

show abstract

Energy Prediction Using Data Analytics in Smart Grid

Anil¹,

Anas²,

Kuruvakkottil³

et al. 2018

ijcse

View full text Add to dashboard Cite

Minimizing the performance metrics of mapreduce workloads

Anusha

Saritha

2017

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

K. Anusha

Performance Evaluation of Spark SQL for Batch Processing

Comparative Study of MongoDB vs Cassandra in big data analytics

Performance Analysis of Apache Spark MLlib Clustering on Batch Data Stored in Cassandra

Energy Prediction Using Data Analytics in Smart Grid

Minimizing the performance metrics of mapreduce workloads

Contact Info

Product

Resources

About