[1]呂宜玲,中文語音辨識中語言模型的強化之研究,國立交通大學資訊工程系所,碩士學位論文,2005。[2]王韋華、徐波,漢語語言模型的規型對統計機器翻譯系統的影響,微計算機信息 Microcomputer Information,2010年第26卷第9-3期。
[3]李民祥、吳世弘、曾議慶、楊秉哲、谷圳,基於對照表以及語言模型之簡繁字體轉換,Computational Linguistics and Chinese Language Processing,Vol. 15,No. 1,March 2010,pp. 19-36。
[4]顧平、朱巧明、李培峰、錢培德,智能型漢字數碼輸入法技術的研究,中文信息學報,2006年第20卷第4期。
[5]賴亦傑,應用多詞及多詞性語言模型的中文斷詞及詞性標記方式,國立中興大學資訊科學與工程學系,碩士學位論文,2011。[6]W. Naptali, Masatoshi Tsuchiya, and Seiichi Nakagawa, Topic-Dependent Language Model with Voting on Noun History, ACM Transactions on Asian Language Information Processing, Vol. 9, No. 2, Article 7 ,2010
[7]王云凱、王萍,基于自然語言處理模型的多音字對漢語拼音字母排序的影響研究,西南民族大學學報自然科學版,2012年第38卷第3期
[8]袁里馳,融合語言知識的統計句法分析,中南大學學報自然科學版,2012年第43卷第3期
[9]陳林、楊丹,獨立于語種的文本分類方法,計算機工程與科學,2008年第30卷第6期
[10]郭雷,統計語言模型分析,軟體導刊,2011年第10卷第11期
[11]Algort P. H. and Cover T. M., 1988, A Sandwich Proof of the Shannon- McMillan-Breiman Theorem, Ahe Annals of Probability, Vol. 16, No. 2, pp. 899-909.
[12]Jurafsky D. and Martin J. H., 2008, Speech and Language Processing (2nd Edition), Prentice Hall, Chapter 6.
[13]袁毓林,基于統計的語言處理模型的局限性,語言文字應用,2004年5月第2期
[14]H. Jeffreys, Theory of Probability, Clarendon Press, Oxford, Second Edition, 1948
[15]Good I. J., 1953, The Population Frequencies of Species and the Estimation of Population Parameters, Biometrika, Vol. 40, pp. 237-264.
[16]Chen Standy F. and Goodman Joshua, 1999, An Empirical Study of Smoothing Techniques for Language Modeling, Computer Speech and Language, Vol. 13, pp. 359-394
[17]Jelinek F., Statistical Methods for Speech Recognition, The MIT Press, Cambridge Massachusetts, 1997.
[18]Nadas A., 1985, On Turing’s Formula for Word Probabilities, IEEE Trans. On Acoustic, Speech and Signal Processing, Vol. ASSP-33, pp. 1414-1416
[19]W. A. Gale and G. Sampson., Good-Turing Frequency Estimation without Tears. Journal of Quantitative Linguistics, 2(3): 15-19, 1995
[20]Ney H. and Essen U., 1991, On Smoothing Techniques for Bigram-Based Natural Language Modeling, IEEE International conference on Acoustic, Speech and Signal Processing, pp. 825-828.
[21]Chinese GigaWord語料庫第三版。
http://www.ldc.upenn.edu/Catalog/CatalogEntry.jsp?catalogId=LDC2007T38
[22]中研院線上斷詞器
http://ckipsvr.iis.sinica.edu.tw/
[23]平衡語料庫簡介
http://db1x.sinica.edu.tw/cgi-bin/kiwi/mkiwi/mkiwi.sh
[24]中央研究院資訊所、語言所詞庫小組所編技術報告第 95-02/98-04號「中央研究院漢語料庫的內容與說明」
[25]Katz S. M., March 1987, Estimation of Probabilities from Sparse Data for the Language Models Component of a Speech Recognizer, IEEE Trans. On Acoustic, Speech and Signal Processing, Vol. ASSP-35, pp. 400-401.