A STORAGE REDUCTION METHOD FOR CORPUS-BASED LANGUAGE MODELS

Hsin-Hsi Chen; Yue-Shi Lee

月旦知識庫會員登入｜元照網路書店｜月旦品評家

熱門：

首頁

臺灣期刊 法律公行政治醫事相關財經社會學教育其他

大陸期刊 核心重要期刊

DOI文章

	本站僅提供期刊文獻檢索。　　【月旦知識庫】是否收錄該篇全文，敬請【登入】查詢為準。最新【購點活動】
篇名	A STORAGE REDUCTION METHOD FOR CORPUS-BASED LANGUAGE MODELS
作者	Hsin-Hsi Chen (Hsin-Hsi Chen)、Yue-Shi Lee (Yue-Shi Lee)
英文摘要	There are many progresses in corpus-based language models recently. However, the storage issue is still one of the major problems in practical applications. This is because the size of the training tables is in direct proportion to the parameters of the language models and the number of the parameters is in direct proportion to the power of these language models. In this paper, we will propose a storage reduction method to solve the problem that results from the large training tables. We use mathematical functions to simulate the distribution of the frequency value of the pairs in the training tables. For the good approximation, the pairs are grouping by their frequency. The experimental results show that although there is a little error rate introduced by the curve function, this scheme is still satisfactory because it performs the closed performance and no extra storage is required in pure curve-fitting model. Besides, we also propose a neural network approach to deal with the pairs classification which is a problem for all class-based approaches. The experimental results show that the neural network approach is suitable to deal with this problem in our storage reduction method.
起訖頁	79-98
刊名	ROCLING論文集
期數	1993 (1993期)
出版單位	國立高雄師範大學輔導與諮商研究所
該期刊-上一篇	Automatic Clustering of Chinese Characters and Words
該期刊-下一篇	A Probabilistic Chunker

新書閱讀

元照讀書館

優惠活動

月旦品評家

元照讀書館

．研討會新訊

月旦知識庫

月旦法律分析庫
月旦醫事法網
月旦會計財稅網

期刊數位服務

社群平台

讀者服務

關於元照

讀者服務專線：+886-2-23756688　傳真：+886-2-23318496
地址：臺北市館前路28 號 7 樓　客服信箱