Home / Current Issue / Paper 1701904
Managing Big Data With Hadoop MapReduce For Solving The Problems Of Traditional RDBMS
Subject area: Science,Engineering and Technology · Area of research: Computer Engineering
Abstract
According to the usage of Internet is extremely increasing, the amount of data being generated by everywhere makes traditional database technologies unable to store and process efficiently and effectively. Moreover, Relational Database Management System (RDBMS) is hardly possible to manage extremely large amount of data called ?Big Data? in structured, semi-structured and unstructured forms from diverse data sources. In this paper, Hadoop, distributed big data processing platform, is applied for managing big data to overcome the storage and processing issues of traditional RDBMS. For experimentation, ?Bag of Words? dataset from UCI Machine Learning Repository is utilized as unstructured big text data tested on Apache Hadoop Map Reduce Multi Node Cluster. According to the outcomes of experimentation, applying Hadoop offers not only faster execution time but also better data scalability performance for managing and processing big text data.
Keywords
Bag of Words, Big Data, Hadoop, Map Reduce, RDBMS
References
[1] B. Gully, H. Eduard, -Extracting data records from unstructured biomedical full text,| in Proc. Joint Conference on Empirical Methods in Natural Language Processing and Computational Natural Language Learning (EMNLP-CoNLL), 2007, pp. 837-846.
[2] B. Peter, B. L. Michael, Christian, -The meaningful use of big data: four perspectives-- four challenges,| ACM Sigmod, vol. 40, pp. 56-60, 2012.
[3] D. Mitchell, M. Neha, S. Vinaya. (2015). Big data vs Traditional data. International Journal for Research in Applied Science & Engineering Technology (IJRASET).
[4] F. George, -Database Trends and Directions: Current Challenges and Opportunities,| in Proc. DATESO, 2010, pp. 163-174.
[5] G. Shankar, R. Siddarth, -Big data analysis using Apache Hadoop,| in Proc. International Conference on IT Convergence and Security (ICITCS), 2014, pp. 1-4.
[6] K. A. Bhadani, J. Dhanya, -Big data: challenges, opportunities, and realities,| in Proc. Effective big data management and opportunities for implementation, 2016, pp. 1-24.
[7] K. Kanimozhi, M. Venkatesan, -Unstructured data analysis-a survey,| International Journal of Advanced Research in Computer and Communication Engineering, vol. 4, pp. 223-225, 2015.
[8] M. Rahul, K. Rashmi, et al., -Efficient analysis of big data using map reduce framework ,| International Journal of Recent Development in Engineering and Technology, vol. 2, 2014.
[9] O. Carlos, -Can we analyze big data inside a DBMS?,| in Proc. 16th international workshop on Data warehousing and OLAP, 2013, pp. 85-92.
[10] P. Jaroslav, -How to Store and Process Big Data: Are Today’s Databases Sufficient?,| in Proc. IFIP International Conference on Computer Information Systems and Industrial Management, Springer, Berlin, 2015, pp. 5-10.
[11] P. Rabi, Padhy, - Big data processing with Hadoop-MapReduce in cloud systems,| International Journal of Cloud Computing and Services Science, vol. 2, pp. 1-16, 2013.
[12] S. Naga, R. Monika, et al., - Data Migration from RDBMS to Hadoop, 2016.
[13] S. Vikram, R. E. Madhusudhan, -Big Data- solutions for RDBMS problems-A survey,| in Proc. 12th IEEE/IFIP Network Operations & Management Symposium, Osaka, Japan, 2013.
[14] T. Abderrahim, B. Abdessamad, et al., -A Comparative Study of Hadoop-based Big Data Architectures,| International Journal of Web Applications (IJWA), vol. 9, pp. 129-137, 2017.
[15] Z. Ahmed, -Data management and big data text analytics,| in Proc. National Conference on Novel Trends in Computer Science (TECHSA-17), 2017, pp. 140-144.
How to cite this paper
@article{1701904,
author = {HTU RA},
title = {Managing Big Data With Hadoop MapReduce For Solving The Problems Of Traditional RDBMS},
journal = {Iconic Research And Engineering Journals},
year = {2020},
volume = {3},
number = {8},
pages = {89-94},
issn = {2456-8880},
url = {https://www.irejournals.com/formatedpaper/1701904.pdf},
abstract = {According to the usage of Internet is extremely increasing, the amount of data being generated by everywhere makes traditional database technologies unable to store and process efficiently and effectively. Moreover, Relational Database Management System (RDBMS) is hardly possible to manage extremely large amount of data called ?Big Data? in structured, semi-structured and unstructured forms from diverse data sources. In this paper, Hadoop, distributed big data processing platform, is applied for managing big data to overcome the storage and processing issues of traditional RDBMS. For experimentation, ?Bag of Words? dataset from UCI Machine Learning Repository is utilized as unstructured big text data tested on Apache Hadoop Map Reduce Multi Node Cluster. According to the outcomes of experimentation, applying Hadoop offers not only faster execution time but also better data scalability performance for managing and processing big text data.},
keywords = {Bag of Words, Big Data, Hadoop, Map Reduce, RDBMS},
month = {February},
}