Home / Current Issue / Paper 1700671
Clustering News Articles For Topic Detection
Subject area: Science,Engineering and Technology · Area of research: Computer engineering
Abstract
In this paper we presented an approach for detecting topics from news articles. Topic detection used in text mining process. Text mining is a field that extract previously unknown and useful information from unstructured textual data. The main purpose of topic detection and tracking is to identify and follow events presented in multiple news sources. Topic detection and tracking would be very helpful to have a system able to map out the data automatically finding story boundaries, determining what stories go with one another, and discovering when something new has happened. We would try to find the first story of new events, identifying all subsequent stories on a certain topic defined by a small number of sample stories, and detect the occurrence of new events. We are going to use agglomeration clustering based on average linkage for detecting the topics.
Keywords
Topic detection, Text Mining, agglomeration clustering.
References
[1] A. A. Kumar, “Text Data Pre-processing and Dimensionality Reduction Techniques for Document Clustering Sri Sivani College of Engineering Sri Sivani College of Engineering,” vol. 1, no. 5, pp. 1–6, 2012.
[2] A. Krause and C. Guestrin, “Data Association for Topic Intensity Tracking,” 2006.
[3] A. Saha, “Learning Evolving and Emerging Topics in Social Media : A Dynamic NMF approach with Temporal Regularization Categories and Subject Descriptors.”
[4] B. Acun, A. Ba, O. Ekin, M. İ. Saraç, and F. Can, “Topic Tracking Using Chronological Term Ranking,” vol. 25, 2011.
[5] B. Pouliquen, R. Steinberger, C. Ignat, E. Käsper, and I. Temnikova, “Multilingual and cross-lingual news topic tracking,” 1998.
[6] C. Elkan, “Text mining and topic models The multinomial distribution,” 2013.
[7] C. Aksoy, F. Can, and S. Kocberber, “Novelty Detection for Topic Tracking,” vol. 63, no. 4, pp. 777–795, 2012.
[8] C. S. Series, R. A-, and A. Xii, Semantic Classes in Topic Detection and Tracking Juha Makkonen. 2009.
[9] C. Cieri, D. Graff, M. Liberman, N. Martey, and S. Strassel, “Large , Multilingual , Broadcast News Corpora For Cooperative Research in Topic Detection And Tracking : The TDT-2 and TDT-3 Corpus Efforts,” no. January 1998, 1999.
[10] D. Eichmann, M. Ruiz, P. Srinivasan, N. Street, C. Culy, and F. Menczer, “A Cluster- Based Approach to Tracking , Detection and Segmentation of Broadcast News.”
[11] F. Perez-tellez, D. Pinto, J. Cardiff, and P. Rosso, “Clustering Weblogs on the Basis of a Topic Detection Method,” pp. 1–10.
[12] F. Fukumoto and Y. Yamaji, “LNAI 3651 - Topic Tracking Based on Linguistic Features,” pp. 10–21, 2005.
[13] I. De, “Experiments in First Story Detection,” pp. 1–8, 2005.
[14] I. Z. B. Bigi, A. Brun, J.P. Haton, K. Smaili, “Dynamic Topic Identification: Towards Combination of Methods.,” pp. 7–9.
[15] J. Allan, S. Harding, D. Fisher, A. Bolivar, S. Guzman-lara, and P. Amstutz, “Taking Topic Detection From Evaluation to Practice,” pp. 1–10, 2004.
[16] J. M. Schultz and M. Liberman, “Topic Detection and Tracking using idf-Weighted Cosine Coefficient,” pp. 2–5.
[17] K. Kaur, “International Journal of Advanced Research in A Survey of Topic Tracking Techniques,” vol. 2, no. 5, pp. 384–393, 2012.
[18] K. Kaur and V. Gupta, “Racking for punjabi language,” vol. 1, no. 3, pp. 37–49, 2011.
[19] K. Megerdoomian and A. Hadjarian, “Automatic Topic Detection in Persian Blogs,” 2011.
[20] M. Mohd, “Design and Evaluation of an Interactive Topic Detection and Tracking Interface,” 2010.
How to cite this paper
@article{1700671,
author = {Vaidehi Patel, Arpita Patel},
title = {Clustering News Articles For Topic Detection},
journal = {Iconic Research And Engineering Journals},
year = {2018},
volume = {1},
number = {11},
pages = {57-61},
issn = {2456-8880},
url = {https://www.irejournals.com/formatedpaper/1700671.pdf},
abstract = {In this paper we presented an approach for detecting topics from news articles. Topic detection used in text mining process. Text mining is a field that extract previously unknown and useful information from unstructured textual data. The main purpose of topic detection and tracking is to identify and follow events presented in multiple news sources. Topic detection and tracking would be very helpful to have a system able to map out the data automatically finding story boundaries, determining what stories go with one another, and discovering when something new has happened. We would try to find the first story of new events, identifying all subsequent stories on a certain topic defined by a small number of sample stories, and detect the occurrence of new events. We are going to use agglomeration clustering based on average linkage for detecting the topics.},
keywords = {Topic detection, Text Mining, agglomeration clustering.},
month = {May},
}