Design of an Input Method for Taiwanese Hokkien using Unsupervized Word Segmentation for Language Modeling

Pierre Magistry

月旦知識庫會員登入｜元照網路書店｜月旦品評家

熱門：

首頁

臺灣期刊 法律公行政治醫事相關財經社會學教育其他

大陸期刊 核心重要期刊

DOI文章

	本站僅提供期刊文獻檢索。　　【月旦知識庫】是否收錄該篇全文，敬請【登入】查詢為準。最新【購點活動】
篇名	Design of an Input Method for Taiwanese Hokkien using Unsupervized Word Segmentation for Language Modeling
並列篇名	Design of an Input Method for Taiwanese Hokkien using Unsupervized Word Segmentation for Language Modeling
作者	Pierre Magistry (Pierre Magistry)
英文摘要	This paper presents the challenges and the methodology followed in the design of a new Input Method (IME) for the Taiwanese (Hokkien) language. We first describe the context, the motivations and some of the main issues related to the input of text in Taiwanese on modern computer systems and mobile devices. Then we present the available resources which our system is based on. We will describe the whole architecture of our system. But since the cornerstone of modern IME is the Language Model (LM), the main Natural Language Processing issue on which we will focus in this paper is the estimation of a LM in the case of this under-resourced language. The solution we propose to rely on unsupervised word segmentation which preserves some degree of ambiguity.
起訖頁	284-298
關鍵詞	Unsupervized Word Segmentation、Language Modeling、Input Method、Taiwanese
刊名	ROCLING論文集
期數	2016 (2016期)
出版單位	中華民國計算語言學學會
該期刊-上一篇	命名實體識別運用於產品同義詞擴增
該期刊-下一篇	Sarcasm Detection in Chinese Using a Crowdsourced Corpus

新書閱讀

元照讀書館

優惠活動

月旦品評家

元照讀書館

．研討會新訊

月旦知識庫

月旦法律分析庫
月旦醫事法網
月旦會計財稅網

期刊數位服務

社群平台

讀者服務

關於元照

讀者服務專線：+886-2-23756688　傳真：+886-2-23318496
地址：臺北市館前路28 號 7 樓　客服信箱