首页    期刊浏览 2025年07月01日 星期二
登录注册

文章基本信息

  • 标题:A Proposed UNICODE-Based Extended Romanization System for Persian Texts
  • 本地全文:下载
  • 作者:M. A. Mahdavi
  • 期刊名称:INTERNATIONAL JOURNAL OF INFORMATION SCIENCE AND MANAGEMENT
  • 印刷版ISSN:2008-8302
  • 电子版ISSN:2008-8310
  • 出版年度:2012
  • 卷号:10
  • 期号:1
  • 页码:57-71
  • 语种:English
  • 出版社:REGIONAL INFORMATION CENTER FOR SCIENCE AND TECHNOLOGY
  • 摘要:So far, various Romanization schemes have been proposed for capturing Persian text using Latin alphabet. However, each have served a very specific and yet limited function. This paper proposes an extended Romanization scheme that can facilitate a wide range of encoding needed in the field of Natural Language Processing. The proposed scheme endeavors to preserve both orthographic and phonological phenomena in the language. It also accounts for encoding handwritten manuscripts, in which glyph ambiguity is a salient feature. It is particularly relevant to Romanizing the Kufi script, in which diacritical marks are omitted. The current work also recommends orthographic rules in an effort to standardize future Romanization tasks.
国家哲学社会科学文献中心版权所有