International Peer-Reviewed JournalOpen AccessISSN 2456-8880
irejournals@gmail.com+91-7433024337

Home / Current Issue / Paper 1700059

1700059 Vol 1 · Issue 3 Download Paper

Supervised Learning Techniques In Web Usage Mining: Comparison And Analysis

Parth Suthar

Subject area: Science,Engineering and Technology  ·  Area of research: Computer Science and Engineering

Abstract

The World Wide Web contains enormous amount of information about almost every imaginable subject. Web mining is a special area of data mining, which deals with identifying interesting patterns and useful information from the web. This knowledge is useful in improving the quality of services provided by web. The information about user access is stored in the form of web access logs at web servers and proxies. Web usage mining is a discipline that deals with extracting users? information regarding user interests and behaviour profiles by processing those web access logs. This knowledge is useful in improving the areas like web personalization, recommendation systems, business intelligence, market segmentation etc. In this paper, we review, analyse and compare various existing supervised learning techniques utilized for web usage mining. We also present methods to compare and test the performance of those techniques

Keywords

supervised learning, web usage mining, web access logs, classification, KNN, SVM, Na?ve Bayes classifier, decision tree classifier, rule based classifier, k-fold cross validation, bootstrapping, confusion matrix

References

[1] Jiawei Han, Michaline Kamber, Jian Pei, Data mining concepts and techniques, 3rd ed., Elsevier Inc., 2012.

[2] [Online]. Available:http://www.slideshare.net/ashrafmat h/naive-bayes-15644818

[3] D.A. Adeniyi, Z. Wei, Y. Yongquan, “Automated web usage data mining and recommendation system using K-Nearest Neighbour (KNN) classification method”, Applied Computing and Informatics, ScienceDirect, Vol. 12, pp. 90-108, 2016.

[4] Thomas G. Dietterich, “Approximate Statistical Test For Comparing The Supervised Classification Learning Algorithms”, December 1997.

[5] Paul Horton , Kenta Nakai, “Better Prediction of Protein Cellular Localization Sites with the Nearest Neighbours Classifier”, American association for Artificial Intelligence, 1997.

[6] Mohammed Hamed Ahmed Elhiber and Ajith Abraham, ”Access Patterns in Web Log Data: A Review”, Journal of Network and Innovative Computing, Vol. 1, pp. 348-355, 2013.

[7] Chintandeep Kaur , Rinkle Rani Aggarwal, “Review on Classification of Web Log Data using CART Algorithm”, International Journal of Computer Applications ,Vol. 80, No. 17 ,October 2013.

[8] D. Jayalatchumy, Dr. P.Thambidurai, “Web Mining Research Issues and Future Directions – A Survey”, IOSR Journal of Computer Engineering (IOSR-JCE), Vol. 14, Issue 03, pp. 20-27, Sep. - Oct. 2013.

[9] K. Sudheer Reddy, G. Partha Saradhi Varma, and M. Kantha Reddy, “An Effective Pre- processing Method for Web Usage Mining”, International Journal of Computer Theory and Engineering, Vol. 06, No. 05, October 2014.

[10] Sujith Jayaprakash, Balamurugan E., “A Comprehensive Survey on Data Preprocessing Methods in Web Usage Mining”, International Journal of Computer Science and Information Technologies, Vol. 06, No 03, pp. 3170-3174, 2015.

[11] Sanjeev Dhawan, Swati Goel, “Web Usage Mining: Finding Usage Patterns from Web Logs”, American International Journal of Research in Science, Technology, Engineering & Mathematics.

[12] Nirali H.Panchal, Ompriya Kale, “A Survey on Web Usage Mining”, International Journal of Computer Trends and Technology (IJCTT), Vol. 17, No. 04, Nov. 2014.

[13] R. Aruna devi, K. Nirmala, “Analysis of Classification Algorithm in Data Mining”, International Journal of Data Mining Techniques and Applications, Vol. 03, pp. 361- 364, Jun 2014.

[14] Jaykumar Jagani, Kamlesh Patel, “An Enhanced Approach for Classification in Web Usage Mining using Neural Network Learning Algorithms for Supervised Learning”, International Journal of Computer Applications, Vol. 90, March 2014.

[15] Ron Kohavi, “A Study of Cross Validation and Bootstrap for Accuracy Estimation and Model Selection”, International Joint Conference on Artificial Intelligence, 1995.

[16] Sofia Visa, Brian Ramsay, Anca Ralescu, Esther van der Knaap, “Confusion Matrix- based Feature Selection”, International Joint Conference on Artificial Intelligence, 1995.

[17] M. Seetha, K. V. N. Sunitha, G. Malini Devi, “Performance Assessment of Neural Network and K-Nearest Neighbour Classification with Random Subwindows”, International Journal of Machine Learning and Computing, Vol. 2, No. 6, December 2012.

[18] Janmenjoy Nayak, Bighnaraj Naik, H. S. Behera, “A Comprehensive Survey on Support Vector Machine in Data Mining Tasks: Applications & Challenges”, International Journal of Database Theory and Application, Vol. 8, No. 1, pp. 169-186, December 2015.

[19] Anshul Bhargav, Munish Bhargav, “Neuro- Fuzzy Based Hybrid Model for Web Usage Mining”, in Eleventh International Multi- Conference on Information Processing, Science Direct, 2015, pp 327 – 334.

[20] A. K. Santra, S. Jayasudha, “Classification of Web Log Data to Identify Interested Users Using Naïve Bayesian Classification”, International Journal of Computer Science Issues, Vol. 9, Issue 1, No 2, January 2012.

[21] Zidrina Pabarskaite, “Decision trees for web log mining”, journal of intelligent data analysis, Vol. 7, Issue 2, April 2003.

[22] A. K. Santra, S. Jayasudha, “Web Page Classification using an Ensemble of Support Vector Machine Classifiers”, Journal Of Networks, Vol. 6, No. 11, November 2011.

How to cite this paper

Parth Suthar "Supervised Learning Techniques In Web Usage Mining: Comparison And Analysis" Iconic Research And Engineering Journals Volume 1 Issue 3 2017 Page 12-18
Parth Suthar "Supervised Learning Techniques In Web Usage Mining: Comparison And Analysis" Iconic Research And Engineering Journals, vol. 1, no. 3, Sep. 2017
Parth Suthar (2017). Supervised Learning Techniques In Web Usage Mining: Comparison And Analysis. Iconic Research And Engineering Journals, 1(3).
Parth Suthar "Supervised Learning Techniques In Web Usage Mining: Comparison And Analysis" Iconic Research And Engineering Journals, vol. 1, no. 3, Sep. 2017.
@article{1700059,
      author = {Parth Suthar},
      title = {Supervised Learning Techniques In Web Usage Mining: Comparison And Analysis},
      journal = {Iconic Research And Engineering Journals},
      year = {2017},
      volume = {1},
      number = {3},
      pages = {12-18},
      issn = {2456-8880},
      url = {https://www.irejournals.com/formatedpaper/1700059.pdf},
      abstract = {The World Wide Web contains enormous amount of information about almost every imaginable subject. Web mining is a special area of data mining, which deals with identifying interesting patterns and useful information from the web. This knowledge is useful in improving the quality of services provided by web. The information about user access is stored in the form of web access logs at web servers and proxies. Web usage mining is a discipline that deals with extracting users? information regarding user interests and behaviour profiles by processing those web access logs. This knowledge is useful in improving the areas like web personalization, recommendation systems, business intelligence, market segmentation etc. In this paper, we review, analyse and compare various existing supervised learning techniques utilized for web usage mining. We also present methods to compare and test the performance of those techniques},
      keywords = {supervised learning, web usage mining, web access logs, classification, KNN, SVM, Na?ve Bayes classifier, decision tree classifier, rule based classifier, k-fold cross validation, bootstrapping, confusion matrix},
      month = {September},
  }