1.《同義詞詞林(擴展版)》(2005),哈爾濱工業大學資訊實驗室提供。
2.《同義詞詞林(擴展版)》說明文件(2005),哈爾濱工業大學資訊實驗室提供。
3.中文斷詞系統網址(中研院),http:// ckipsvr.iis.sinica.edu.tw/.
4.陳信希(2000),「自動摘要方法之研究:單一中文文本之摘要」,行政院國家科學委員會研究計畫,計劃編號:NSC89-2213-E002-064。
5.劉群、李素建(2002),「基於《知網》的詞彙語義相似度計算」,第三屆漢語詞彙語義學研討會論文集,臺北:,pp. 59-76。
6.蕭文峰、張德民、胡國信,「以遞增式分群為基之分類方法過濾具偏斜類別及概念漂移之垃圾郵件」,資訊管理學報(已接受,2007/12)。7.蕭文峰、劉凱帆,2006,「以自動摘要為基礎之中文文件分類器」,第十七屆國際資訊管理學術研討會論文集,義守大學,高雄。
8.Aas, K. and Eikvil, L. (1999), “Text Categorization: A Survey,” Technical report, Norwegian Computing Center, Junho
9.Angheluta, R., De Busser, R., & Moens, M.F. (2002). “The use of topic segmentation for automatic summarization,” In U. Hahn & D. Harman (Eds.), Proceedings of the workshop on automatic summarization, Philadelphia, Pennsylvania, USA, pp. 66-70.
10.Baker, L.D. and Mccallum, A.K. (1998), "Distributional clustering of words for text classification,” In Proceedings of the 21th Ann Int ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR’98), pp. 96-103.
11.Brunn, M., Chali, Y., Pinchak, C. J. (2001). “Text Summarization Using Lexical Chains,” Workshop on Text Summarization, ACM SIGIR Conference. New Orleans, Louisiana USA.
12.Chen, H.H., Kuo, J.J., Huang, S.J., Lin, C.J., and Wung, H.C. (2003), "A Summarization System for Chinese News from Multiple Sources," Journal of American Society for Information Science and Technology, 54(13), November 2003, pp. 1224-1236.
13.Chen, K.H. (1995), “Topic Identification in Discourse,” In Proceedings of the 7th Conference of the European Chapter of Association for Computational Linguistics, pp. 267-271, Dublin, Ireland
14.Chen, K.H., Chen, H.H. (1995), “A Corpus-Based Approach to Text Partition,” In Proceedings of the Workshop of Recent Advances in Natural Language Processing , pp. 152-161, Sofia, Bulgaria
15.Chen, K.H., Huang, S.J., Lin, W.C., and Chen, H.H. (1998), “An NTU-Approach to Automatic Sentence Extraction for Summary Generation,” In Proceedings of the First Automatic Text Summarization Conference (SUMMAC-1), pp. 163-170, Virginia, May
16.Ferrier, L. (2001), A Maximum Entropy Approach to Text Summarization, School of Artificial Intelligence, Division of Informatics, University of Edinburgh.
17.Goldstein, J., Kantrowitz, M., Mittal, V., and Carbonell, J. (1999), “Summarizing Text Documents: Sentence Selection and Evaluation Metrics,” In Proceedings of ACM-SIGIR'99, Berkeley, CA
18.Hearst, M.A. (1999), “Untangling text data mining,” Proceedings of ACL’99: the 37th annual meeting of the association for computational linguistics, University of Maryland.
19.Hsiao, W.F. and Chang, T.M. (2007), “An Incremental Cluster-based Approach to SPAM Filtering,” Expert Systems with Applications (in press, Available online 28 January 2007)
20.Hand, T.F., Sundheim, B. (1998). "TIPSTER-SUMMAC Summarization Evaluation," Proceedings of the TIPSTER Text Phase III Workshop, Washington DC, USA, pp. 353-340.
21.Inderjit S., D., Subramanyam, M., and Rahul, K. (2003), “A Divisive Information-Theoretic Feature Clustering Algorithm for Text Classification”
22.Joachims, T. (1998), “Text categorization with support vector machines: learning with many relevant features,” In Proceedings of ECML-98, 10th European Conference on Machine Learning (Chemnitz, Germany, 1998), pp. 137–142.
23.Kan, M.Y., Klavans, J. (2002), "Using librarian techniques in automatic text summarization for information retrieval," Proceedings of the 2nd ACM/IEEE-CS joint conference on Digital libraries, pp.36-45.
24.Katz S., M. (1995), “Distribution of content words and phrases in text and language modeling,” Natural Language Engineering, Vol. 2(1), pp. 15–59.
25.K Nearest Neighbor Classifier (2008), http://en.wikipedia.org/wiki/Nearest_neighbor_ (pattern_ recognition)
26.Lang, K. (1995), NEWSWEEDER: learning to filter netnews. In Proceedings of ICML-95, 12th International Conference on Machine Learning (Lake Tahoe, CA, 1995), 331-339.
27.Li, S.J., Zhang, J., Huang, X., Bai, S. and Liu, Q. (2002), “Semantic Computation in a Chinese Question-Answering System,” Journal of Computer Science & Technology(JCST), vol.17, No.6, pp. 933 – 939.
28.Liang, C.Y., Guo, L., Xia, Z.J., Nie, F.G., Li, X.X., Su, L., and Yang, Z.Y. (2006), “Dictionary-based text categorization of chemical web pages,” Information Processing and Management Volume: 42, Issue: 4.
29.Luhn, H.P. (1958), “The automatic creation of literature abstracts,” I.B.M. Journal of Research and Development, 2 (2), pp. 159-165.
30.Lingpipe, http://alias-i.com/lingpipe/index.html
31.Ma, L.P., Shepherd, J. and Zhang, Y.C.h. (2003), “Enhancing text classification using synopses extraction,” In Proceeding of the fourth international conference on web information systems engineering, pp. 115–124.
32.Nakao, Y. (2000), An Alogrithm for One-page Summarization of a Long Text Based on Thematic Hierarchy Detection, Fujitsu Laboratories Ltd.
33.Nigam, K., Mccallum, A.K., Thrun, S., and Mitchell, T. (2000), “Text Classification from Labeled and Unlabeled Documents using EM,” Machine Learning, Vol. 39, pp. 103-134.
34.Noam, S. and Naftali, T. (2001), “Agglomerative Information Bottleneck”
35.Naïve Bayes Classifier, http://nlp.stanford.edu/IR-book/html/htmledition/naive-bayes- text -classification-1.html
36.Porter, M. (1980), “An algorithm for suffix stripping,” Automated Library and Information Systems, Vol. 14, No. 3, pp. 130-137
37.Saravanan, M. and Raman, S. (2002), “The term distribution model for summarization of multiple documents,” In Proc. of the Indo European Conference on Multilingual Communication Technologies (IEMCT 2002), pp. 182–192.
38.Saravanan, M., Reghuraj P., C., and Raman, S. (2003), “Summarization and categorization of text data in high-level data cleaning for information retrieval,” Applied Artificial Intelligence, Vol. 17, pp. 461–474.
39.Sebastiani, F. (2002), “Machine learning in automated text categorization,” ACM Computing Surveys, Vol. 34, No.1, pp.1-47.
40.Seidl, T. and Kriegel, H. (1998), “Optimal Multi-Step k-Nearest neighbor search,” Proceedings of ACM SIGMOD Internet Conference On Management of Data, pp. 154-165.
41.Silla Jr., C. N., Kaestner, C. A. A., and Freitas, A. A. (2003), “A non-linear topic detection method for text summarization using WordNet,” http://citeseer.ist.psu.edu/703670.html.
42.Spärck Jones, K. (1999), “Automatic Summarizing: Factors and Directions,” in I. Mani and M. Maybury (eds.), Advances in Automatic Text Summarization, MIT Press, Cambridge, MA.
43.Spärck Jones, K. (2007), “Automatic summarising: The state of the art,” Information Processing and Management, Vol. 43(6), pp. 1449-1481.
44.Tsay, J.J. and Wang, J.D. (2000), “Design and Evaluation of Approaches to Automatic Chinese Text Categorization,” International Journal of Computational Linguistics & Chinese Language Processing, Vol. 5, No. 2, pp. 43-58.
45.Toutanova, K., Klein, D., Manning, C., and Singer, Y. (2003), “Feature-Rich Part-of-Speech Tagging with a Cyclic Dependency Network,” Proceedings of Human Language Technology Conference of the North American Chapter of the Association for Computational Linguistics (HLT-NAACL 2003), pp. 252-259
46.Wang, Z.W., Wong, S.K.M. and Yao, Y.Y. (1992),”An Analysis of Vector Space Models Based on Computational Geometry,” Proceedings of the 15th Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, pp. 152-160.
47.Wei, C., Hu, P., Huang, C.N., Yang, C.S., and Tai, C.H. (2004), “Managing Word Mismatch Challenge in Information Retrieval: A Clustering-Based Query Expansion Method,” Proceedings of the Third Workshop on e-Business (WEB 2004), Washington D.C., pp.82-92.
48.Witten, I.H. and Frank, E. (2005), Data Mining: Practical Machine Learning Tools and Techniques, 2nd Edition, Morgan Kaufmann Series in Data Management Systems.
49.Yang, Y.M. and Liu, X. (1999), "A re-examination of text categorization methods," Proceedings of ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR'99), pp. 42-49.