Wikipedia Mining

Kotaro NAKAYAMA, Masahiro ITO, Maike ERDMANN, Masumi SHIRAKAWA, Tomoyuki MICHISHITA, Takahiro HARA, Shojiro NISHIO
2009 Transactions of the Japanese society for artificial intelligence  
Wikipedia, a collaborative Wiki-based encyclopedia, has become a huge phenomenon among Internet users. It covers a huge number of concepts of various fields such as arts, geography, history, science, sports and games. As a corpus for knowledge extraction, Wikipedia's impressive characteristics are not limited to the scale, but also include the dense link structure, URL based word sense disambiguation, and brief anchor texts. Because of these characteristics, Wikipedia has become a promising
more » ... us and a new frontier for research. In the past few years, a considerable number of researches have been conducted in various areas such as semantic relatedness measurement, bilingual dictionary construction, and ontology construction. Extracting machine understandable knowledge from Wikipedia to enhance the intelligence on computational systems is the main goal of "Wikipedia Mining," a project on CREP (Challenge for Realizing Early Profits) in JSAI. In this paper, we take a comprehensive, panoramic view of Wikipedia Mining research and the current status of our challenge. After that, we will discuss about the future vision of this challenge.
doi:10.1527/tjsai.24.549 fatcat:o7e5grmn5rgctf67y7lhfppq6i