文章基本信息

标题：Statistical string similarity model for information linkage
作者：Atsuhiro TAKASU
期刊名称：Progress in Informatics
印刷版ISSN：1349-8614
电子版ISSN：1349-8606
出版年度：2009
期号：6
页码：57-62
DOI：10.2201/NiiPi.2009.6.7
出版社：National Institute of Informatics
摘要：This paper proposes a statistical string similarity model for approximate matching in information linkage. The proposed similarity model is an extension of hidden Markov model and its learnable ability realizes string matching function adaptable to various information sources. The main contribution of this paper is to develop an efficient learning algorithm for estimating parameters of the statistical similarity model. The proposed algorithm is based on the Expectation-Maximization (EM) technique where dynamic programing technique is used to update parameters in EM process.
关键词：String similarity; statistical model; EM algorithm