文章基本信息

标题：Wikipedia Mining Challenge for Realizing Early Profits, The Kick Off
本地全文：下载
作者：Kotaro NAKAYAMA ; Masahiro ITO ; Maike ERDMANN 等
期刊名称：人工知能学会論文誌
印刷版ISSN：1346-0714
电子版ISSN：1346-8030
出版年度：2009
卷号：24
期号：6
页码：549-557
DOI：10.1527/tjsai.24.549
出版社：The Japanese Society for Artificial Intelligence
摘要：Wikipedia, a collaborative Wiki-based encyclopedia, has become a huge phenomenon among Internet users. It covers a huge number of concepts of various fields such as arts, geography, history, science, sports and games. As a corpus for knowledge extraction, Wikipedia's impressive characteristics are not limited to the scale, but also include the dense link structure, URL based word sense disambiguation, and brief anchor texts. Because of these characteristics, Wikipedia has become a promising corpus and a new frontier for research. In the past few years, a considerable number of researches have been conducted in various areas such as semantic relatedness measurement, bilingual dictionary construction, and ontology construction. Extracting machine understandable knowledge from Wikipedia to enhance the intelligence on computational systems is the main goal of "Wikipedia Mining," a project on CREP (Challenge for Realizing Early Profits) in JSAI. In this paper, we take a comprehensive, panoramic view of Wikipedia Mining research and the current status of our challenge. After that, we will discuss about the future vision of this challenge.
关键词：Wikipedia Mining ; Social Media ; Ontology ; Thesaurus