Transactions of the Japanese Society for Artificial Intelligence
Online ISSN : 1346-8030
Print ISSN : 1346-0714
ISSN-L : 1346-0714
Original Paper
Wikipedia Mining
Challenge for Realizing Early Profits, The Kick Off
Kotaro NAKAYAMAMasahiro ITOMaike ERDMANNMasumi SHIRAKAWATomoyuki MICHISHITATakahiro HARAShojiro NISHIO
Author information
JOURNAL FREE ACCESS

2009 Volume 24 Issue 6 Pages 549-557

Details
Abstract

Wikipedia, a collaborative Wiki-based encyclopedia, has become a huge phenomenon among Internet users. It covers a huge number of concepts of various fields such as arts, geography, history, science, sports and games. As a corpus for knowledge extraction, Wikipedia's impressive characteristics are not limited to the scale, but also include the dense link structure, URL based word sense disambiguation, and brief anchor texts. Because of these characteristics, Wikipedia has become a promising corpus and a new frontier for research. In the past few years, a considerable number of researches have been conducted in various areas such as semantic relatedness measurement, bilingual dictionary construction, and ontology construction. Extracting machine understandable knowledge from Wikipedia to enhance the intelligence on computational systems is the main goal of "Wikipedia Mining," a project on CREP (Challenge for Realizing Early Profits) in JSAI. In this paper, we take a comprehensive, panoramic view of Wikipedia Mining research and the current status of our challenge. After that, we will discuss about the future vision of this challenge.

Content from these authors
© 2009 JSAI (The Japanese Society for Artificial Intelligence)
Previous article Next article
feedback
Top