首页    期刊浏览 2024年09月20日 星期五
登录注册

文章基本信息

  • 标题:Applications of Lexicographic Semirings to Problems in Speech and Language Processing
  • 本地全文:下载
  • 作者:Richard Sproat ; Mahsa Yarmohammadi ; Izhak Shafran
  • 期刊名称:Computational Linguistics
  • 印刷版ISSN:0891-2017
  • 电子版ISSN:1530-9312
  • 出版年度:2014
  • 卷号:40
  • 期号:4
  • 页码:733-761
  • DOI:10.1162/COLI_a_00198
  • 语种:English
  • 出版社:MIT Press
  • 摘要:This paper explores lexicographic semirings and their application to problems in speech and language processing. Specifically, we present two instantiations of binary lexicographic semirings, one involving a pair of tropical weights, and the other a tropical weight paired with a novel string semiring we term the categorial semiring . The first of these is used to yield an exact encoding of backoff models with epsilon transitions. This lexicographic language model semiring allows for off-line optimization of exact models represented as large weighted finite-state transducers in contrast to implicit (on-line) failure transition representations. We present empirical results demonstrating that, even in simple intersection scenarios amenable to the use of failure transitions, the use of the more powerful lexicographic semiring is competitive in terms of time of intersection. The second of these lexicographic semirings is applied to the problem of extracting, from a lattice of word sequences tagged for part of speech, only the single best-scoring part of speech tagging for each word sequence. We do this by incorporating the tags as a categorial weight in the second component of a 〈Tropical, Categorial〉 lexicographic semiring, determinizing the resulting word lattice acceptor in that semiring, and then mapping the tags back as output labels of the word lattice transducer. We compare our approach to a competing method due to Povey et al. (2012).
国家哲学社会科学文献中心版权所有