月旦知識庫
 
  1. 熱門:
 
首頁 臺灣期刊   法律   公行政治   醫事相關   財經   社會學   教育   其他 大陸期刊   核心   重要期刊 DOI文章
中文計算語言學期刊 本站僅提供期刊文獻檢索。
  【月旦知識庫】是否收錄該篇全文,敬請【登入】查詢為準。
最新【購點活動】


篇名
Analyzing the Morphological Structures in Seediq Words
作者 Chuan-Jie Lin (Chuan-Jie Lin)Li-May SungJing-Sheng YouWei Wang (Wei Wang)Cheng-Hsun LeeZih-Cyuan Liao
英文摘要
NLP techniques are efficient to build large datasets for low-resource languages. It is helpful for preservation and revitalization of the indigenous languages. This paper proposes approaches to analyze morphological structures in Seediq words automatically as the first step to develop NLP applications such as machine translation. Word inflections in Seediq are plentiful. Sets of morphological rules have been created according to the linguisitic features provided in the Seediq syntax book (Sung, 2018) and based on regular morpho-phonological processing in Seediq, a new idea of "deep root" is also suggested. The rule-based system proposed in this paper can successfully detect the existence of infixes and suffixes in Seediq with a precision of 98.88% and a recall of 89.59%. The structure of a prefix string is predicted by probabilistic models. We conclude that the best system is bigram model with back-off approach and Lidstone smoothing with an accuracy of 82.86%.
起訖頁 1-20
關鍵詞 SeediqAutomatic Analysis of Morphological StructuresDeep RootNatural Language Processing for Indigenous Languages in TaiwanFormosan Languages
刊名 中文計算語言學期刊  
期數 202012 (25:2期)
出版單位 中華民國計算語言學學會
該期刊-下一篇 基於圖神經網路之中文健康照護命名實體辨識
 

新書閱讀



最新影音


優惠活動




讀者服務專線:+886-2-23756688 傳真:+886-2-23318496
地址:臺北市館前路28 號 7 樓 客服信箱
Copyright © 元照出版 All rights reserved. 版權所有,禁止轉貼節錄