月旦知識庫
 
  1. 熱門:
 
首頁 臺灣期刊   法律   公行政治   醫事相關   財經   社會學   教育   其他 大陸期刊   核心   重要期刊 DOI文章
中文計算語言學期刊 本站僅提供期刊文獻檢索。
  【月旦知識庫】是否收錄該篇全文,敬請【登入】查詢為準。
最新【購點活動】


篇名
An Assessment of Character-based Chinese News Filtering Using Latent Semantic Indexing
作者 Wu, Shih-hung (Wu, Shih-hung)Yang, Pey-ching (Yang, Pey-ching)Soo,Von-wun (Soo,Von-wun)
中文摘要
We assess the Latent Semantic Indexing (LSI) approach to Chinese information filtering. In particular, the approach is for Chinese news filtering agents that use a character-based and hierarchical filtering scheme. The traditional vector space model is employed as an information filtering model, and each document is converted into a vector of weights of terms. Instead of using words as terms in the IR nominating tradition, terms refer to Chinese characters. LSI captures the semantic relationship between documents and Chinese characters. We use the Sin-gular-value Decomposition (SVD) technique to compress the term space into a lower dimension which achieves latent association between documents and terms. The results of experiments show that the recall and precision rates of Chinese news filtering using the character-based ap-proach incorporating the LSI technique are satisfactory.
起訖頁 61-78
關鍵詞 中文資訊檢索中文資訊過濾
刊名 中文計算語言學期刊  
期數 199808 (3:2期)
出版單位 中華民國計算語言學學會
該期刊-上一篇 Information Extraction: Beyond Document Retrieval
該期刊-下一篇 Noisy Channel Models for Corrupted Chinese Text Restoration and GB-to-Big5 Conversion
 

新書閱讀



最新影音


優惠活動




讀者服務專線:+886-2-23756688 傳真:+886-2-23318496
地址:臺北市館前路28 號 7 樓 客服信箱
Copyright © 元照出版 All rights reserved. 版權所有,禁止轉貼節錄