跳到主要內容

臺灣博碩士論文加值系統

(216.73.216.221) 您好!臺灣時間:2026/10/05 04:56
字體大小: 字級放大   字級縮小   預設字形  
回查詢結果 :::

詳目顯示

我願授權國圖
: 
twitterline
研究生:洪慶霖
研究生(外文):Ching-Lin Hung
論文名稱:事件與意見摘要方法之研究
論文名稱(外文):A Study of Event and Opinion Summarization
指導教授:陳信希陳信希引用關係
指導教授(外文):Hsin-Hsi Chen
學位類別:碩士
校院名稱:國立臺灣大學
系所名稱:資訊工程學研究所
學門:工程學門
學類:電資工程學類
論文種類:學術論文
論文出版年:2007
畢業學年度:95
語文別:中文
論文頁數:81
中文關鍵詞:摘要
外文關鍵詞:Summarization
相關次數:
  • 被引用被引用:0
  • 點閱點閱:264
  • 評分評分:
  • 下載下載:0
  • 收藏至我的研究室書目清單書目收藏:0
隨著網際網路的蓬勃發展,文字資訊逐漸電子化,在大量的資料下,如何有效率的瀏覽感到興趣的資訊變得越來越重要。在過去的相關研究中,為了可以有效率的瀏覽資訊,而提出多文件摘要。
然而,傳統的多文件摘要著重於文件重要內容的整理分析,但是,若使用者有興趣的資訊是社會大眾對某個事件的看法,而不僅僅是事件內容的話,這樣的多文件摘要並不能滿足使用者的需求。根據觀察,使用者有興趣的資訊愈來愈傾向含有意見的資訊,而不再單單是事件本身。
本篇論文著重在簡短的摘要,跟傳統採用壓縮比的方式不同,我們所產生的摘要約在400字的大小。現在資料的取得越來越容易,要取得大量的文件不是困難的事情,但要在短時間內知道大量文件中的內容是很困難的,如果採用傳統壓縮比的方式去產生摘要,資料量很大時還是要花不少時間去閱讀摘要,所以我們想在短短400字摘要裡面盡量包含文件的重要內容,讓讀者可以在短時間內了解到文件中所要表達的事情。
本篇論文提出,用事件摘要跟意見摘要經過適當組合產生綜合摘要,事件摘要用事件詞分類,來將討論到相似事件的句子放在同一群,以避免選到內容相似的句子,再用事件分數去挑出代表各群的句子。為了產生意見摘要,我們找出一群對意見有鑑別力的詞性來判斷句子是否為意見句,利用這些詞性來算出句子的意見分數,並根據意見分數選出具有比較強烈的意見的句子。最後將產生的事件跟意見摘要,經過最適當的組合產生最後的綜合摘要。
摘要 i
索引 iii
附圖目錄 v
附表目錄 iv
第1章 緒論 1
1.1 研究動機 1
1.2 事件摘要介紹 1
1.3 意見摘要介紹 2
1.4 綜合摘要介紹 2
1.5 相關研究 2
1.6 實驗文件集介紹 3
1.7 系統架構 8
1.8 論文編排 9
第2章 事件摘要 10
2.1 目的 10
2.2 事件偵測 10
2.2.1 目的 10
2.2.2 特徵選取 10
2.3 事件摘要方法 13
2.3.1 分群 14
2.3.1.1 句子事件詞向量表示法 14
2.3.1.2 分群演算法 14
2.3.2 選群 15
2.3.3 選句 16
2.4 事件摘要實驗與討論 17
2.4.1 選群 17
2.4.2 選句 17
2.4.3 摘要評估 18
2.5 事件摘要呈現 21
第3章 意見摘要 27
3.1 目的 27
3.2 對意見句有鑑別力的詞性 27
3.2.1 如何找出對意見有鑑別力的詞性 27
3.2.2 找出對意見有鑑別力的詞性流程圖一 29
3.2.3 NTCIR答案介紹 30
3.2.4 找出對意見有鑑別力的詞性流程圖二 31
3.2.5 找出對意見有鑑別力的詞性實驗結果 32
3.2.6 用詞性找意見句的評估 37
3.3 詞性分數 38
3.3.1 詞性規則 38
3.3.2 詞性分數評估 39
3.4 意見摘要方法 43
3.4.1 分群 43
3.4.2 選群 44
3.4.3 選句 44
3.5 意見摘要實驗與討論 44
3.5.1 選群 44
3.5.2 選句 45
3.5.3 摘要評估 46
3.6 意見摘要呈現 48
第4章 綜合摘要 50
4.1 目的 50
4.2 綜合摘要方法 50
4.2.1 事件句與意見句在綜合摘要中的比例 50
4.2.2 選句 50
4.3 綜合摘要實驗與討論 51
4.4 主題明確 55
4.5 事件的串連性 57
4.6 應用正規分布調整句子分數 59
4.7 綜合摘要呈現 69
第5章 簡短摘要與一般摘要 70
5.1 簡短摘要與一般摘要的比較 70
5.2 應用摘要方法於一般摘要之實驗與討論 70
5.3 摘要呈現 71
第6章 討論與未來研究 76
6.1 結論 76
6.2 討論 77
6.3 未來研究 77
第7章 參考資料 79
[1]李俐瑩,意見摘要方法之研究,碩士論文,2005
[2]Brunn, M., Chali, Y., & Pinchak, C.J. (2001). Text summarization using lexical chains. In Proceedings of First Document Understanding Conference, New Orleans, LA.
[3]Edmundson, H.P. (1964). Problems in automatic extracting. Communications of the ACM, 7, pp.59–263.
[4]Edmundson, H.P. (1969). New methods in automatic extracting. Journal of the ACM, 16, pp. 264–285. 2.Hovy, E., & Marcu, D. (1998). Automated text summarization. Tutorial in COLING/ACL98.
[5]Kupiec, J.M., Peterson, J., & Chen, F. (1995). A trainable document summarizer. Proceeding of the 18th Annual International Conference on Research and Development in Information Retrieval (ACM SIGIR ’95), pp. 68–73, Seattle, WA.
[6] Kuo, J.-J. and Chen H.-H. Cross-document event clustering using knowledge mining from co-reference chains, IP&M, 2007
[7]Lin, C.Y., & Hovy, E. (1997). Identifying topics by position. In Proceedings of the 5th ACL Conference on Applied Natural Language Processing, pp. 283–290, Washington, DC.
[8] Dave, K., Lawrence, S. and Pennock, D.M. Mining the Peanut Gallery: Opinion Extraction and Semantic Classification of Product Reviews. Proceedings of 12th International Conference on World Wide Web, pages 519-528, 2003.
[9] Fukumoto, Fumiyo and Suzuki, Yoshimi. Event Tracking based on Domain Dependency. Proceedings of the 23rd Annual International ACM SIGIR Conference on Research and Development in Information Retrieval, pages 57-64, 2000.
[10] Hu, M. and Liu, B. Mining and Summarizing Customer Reviews. Proceedings of the ACM SIGKDD International Conference on Knowledge Discovery & Data Mining, Seattle, Washington, USA, Aug 22-25, 2004.
[11] Hu, M. and Liu, B. Mining Opinion Features in Customer Reviews. Proceedings of Nineteeth National Conference on Artificial Intelligence, San Jose, USA, July 2004.
[12]Knight, K. & Marcu, D. (2000). Statistics-based summarization – step one: Sentence compression. In Proceedings of AAAI-2000, Austin, TX.
[13]Kuo, J.-J. and Chen H.-H. Event Clustering on Streaming News Using Co-Reference Chains and Event Words. Proceedings of ACL 2004 Workshop on Reference Resolution and Its Applications, July 25-26, Barcelona, Spain, pages 17-23, 2004.
[14] Ku et al., Construction of an Evaluation Corpus for Opinion Extraction,2005
[15]Liu, B., Hu, M. and Cheng, J. Opinion Observer: Analyzing and Comparing Opinions on the Web. Proceedings of the 14th international World Wide Web conference, May 10-14, in Chiba, Japan, pages 342-351, 2005.
[16]Luhn, H.P. (1958). The automatic creation of literature abstract. IBM Journal of Research and Development, 2(2), pp. 159–165.
[17]Mani, I., & Bloedorn, E. (1997). Multi-document summarization by graph search and matching. In Proceedings of the 14th National Conference on Artificial Intelligence, pp. 622–628, Providence, RI.
[18]Miller, G.-A., Beckwith, R., Fellbaum, C., Gross, D. and Miller, K. Introduction to WordNet: An On-line Lexical Database. Journal of Lexicography, 3(4), pages 235-244, 1990.
[19] Morinaga, S., Yamanishi, K., Tateishi, K. and Fukushima, T. Mining Product Reputations on the Web. Proceedings of the Eighth International Conference on Knowledge Discovery and Data Mining, July 23-26, pages 341-349, 2002.
[20] Huang S.-J. and Chen H.-H. A Summarization System for Chinese News from Multiple Sources. Proceedings of the 4th International Workshop on Information Retrieval with Asian Languages, November 11-12, 1999.
[21]McKeown, K.R., & Radev, D.R. (1995). Generating summaries of multiple news articles. In Proceedings of 18th Annual International Conference on Research and Development in Information Retrieval (ACM SIGIR ‘95), pp. 74–82, Seattle, WA.
[22]McKeown, K.R., Klavans, J.L., Hatzivassiloglou V., Barzilay, R., & Eskin, E. (1999). Toward multidocument summarization by reformulation: Progress and prospects. In Proceedings of the Seventeenth National Conference on Artificial Intelligence (AAAI-99), pp. 453–460, Orlando, FL.
[23] Radev, D.R., Blair-Goldensohn, S., & Zhang, Z. (2001). Experiments in single and multi-document summarization using MEAD. In Proceedings of first document understanding conference, New Orleans, LA.
[24]Radev, D.R., Jing, H.Y., Stys, M. & Tam, D. (2004). Centroid-based summarization of multiple documents. Journal of Information Processing and Management 40, pp.919-958.
QRCODE
 
 
 
 
 
                                                                                                                                                                                                                                                                                                                                                                                                               
第一頁 上一頁 下一頁 最後一頁 top