Beyond SumBasic: Task-focused summarization with sentence simplification and lexical expansion
作者:
Highlights:
•
摘要
In recent years, there has been increased interest in topic-focused multi-document summarization. In this task, automatic summaries are produced in response to a specific information request, or topic, stated by the user. The system we have designed to accomplish this task comprises four main components: a generic extractive summarization system, a topic-focusing component, sentence simplification, and lexical expansion of topic words. This paper details each of these components, together with experiments designed to quantify their individual contributions. We include an analysis of our results on two large datasets commonly used to evaluate task-focused summarization, the DUC2005 and DUC2006 datasets, using automatic metrics. Additionally, we include an analysis of our results on the DUC2006 task according to human evaluation metrics. In the human evaluation of system summaries compared to human summaries, i.e., the Pyramid method, our system ranked first out of 22 systems in terms of overall mean Pyramid score; and in the human evaluation of summary responsiveness to the topic, our system ranked third out of 35 systems.
论文关键词:Summarization,Multi-document summarization,Sentence simplification,Lexical expansion,Query expansion,NLP
论文评审过程:Received 18 July 2006, Revised 18 January 2007, Accepted 22 January 2007, Available online 19 April 2007.
论文官网地址:https://doi.org/10.1016/j.ipm.2007.01.023