Next Article in Journal
Handling Semantic Complexity of Big Data using Machine Learning and RDF Ontology Model
Previous Article in Journal
Optimal Decision in a Dual-Channel Supply Chain under Potential Information Leakage
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Research Front Detection and Topic Evolution Based on Topological Structure and the PageRank Algorithm

School of Information, Zhejiang University of Finance and Economics, Hangzhou 310018, China
*
Author to whom correspondence should be addressed.
Symmetry 2019, 11(3), 310; https://doi.org/10.3390/sym11030310
Submission received: 8 January 2019 / Revised: 22 February 2019 / Accepted: 24 February 2019 / Published: 1 March 2019

Abstract

Research front detection and topic evolution has for a long time been an important direction for research in the informetrics field. However, most previous studies either simply use a citation count for scientific document clustering or assume that each scientific document has the same importance in detecting the clustering theme in a cluster. In this study, utilizing the topological structure and the PageRank algorithm, we propose a new research front detection and topic evolution approach based on graph theory. This approach is made up of three stages: (1) Setting a time window with appropriate length according to the accuracy of scientific documents clustering results and the time delay of a scientific document to be cited, dividing scientific documents into several time windows according to their years of publication, calculating similarities between them according to their topological structure, and clustering them in each time window based on the fast greedy algorithm; (2) combining the PageRank algorithm and keywords’ frequency to detect the clustering theme, which assumes that the more important a scientific document in the cluster is, the greater the possibility that it is cited by the other documents in the same cluster; and (3) reconstructing the cluster graph where nodes represent clusters and edges’ strengths represent the similarities between different clusters, then detecting research front and identifying topic evolution based on the reconstructed cluster graph. To evaluate the performance of our proposed approach, the scientific documents related to data mining and covered by Science Citation Index Expanded (SCI-EXPANDED) or Social Science Citation Index (SSCI) in Web of Science are collected as a case study. The experiment’s results show that the proposed approach can obtain reasonable clustering results, and it is effective for research front detection and topic evolution.
Keywords: research front detection; topic evolution; topological structure; PageRank algorithm; fast greedy algorithm; keywords frequency research front detection; topic evolution; topological structure; PageRank algorithm; fast greedy algorithm; keywords frequency

Share and Cite

MDPI and ACS Style

Xu, Y.; Zhang, S.; Zhang, W.; Yang, S.; Shen, Y. Research Front Detection and Topic Evolution Based on Topological Structure and the PageRank Algorithm. Symmetry 2019, 11, 310. https://doi.org/10.3390/sym11030310

AMA Style

Xu Y, Zhang S, Zhang W, Yang S, Shen Y. Research Front Detection and Topic Evolution Based on Topological Structure and the PageRank Algorithm. Symmetry. 2019; 11(3):310. https://doi.org/10.3390/sym11030310

Chicago/Turabian Style

Xu, Yangbing, Shuai Zhang, Wenyu Zhang, Shuiqing Yang, and Yue Shen. 2019. "Research Front Detection and Topic Evolution Based on Topological Structure and the PageRank Algorithm" Symmetry 11, no. 3: 310. https://doi.org/10.3390/sym11030310

APA Style

Xu, Y., Zhang, S., Zhang, W., Yang, S., & Shen, Y. (2019). Research Front Detection and Topic Evolution Based on Topological Structure and the PageRank Algorithm. Symmetry, 11(3), 310. https://doi.org/10.3390/sym11030310

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop