Presentations (Communicative Events)

Towards multidocument summarization by reformulation: progress and prospects

McKeown, Kathleen; Klavans, Judith L.; Hatzivassiloglou, Vasileios; Barzilay, Regina; Eskin, Eleazar

By synthesizing information common to retrieved documents, multi-document summarization can help users of information retrieval systems to find relevant documents with a minimal amount of reading. We are developing a multidocument summarization system to automatically generate a concise summary by identifying and synthesizing similarities across a set of related documents. Our approach is unique in its integration of machine learning and statistical techniques to identify similar paragraphs, intersection of similar phrases within paragraphs, and language generation to reformulate the wording of the summary. Our evaluation of system components shows that learning over multiple extracted linguistic features is more effective than information retrieval approaches at identifying similar text units for summarization and that it is possible to generate a fluent summary that conveys similarities among documents even when full semantic interpretations of the input text are not available.

Files

More About This Work

Academic Units
Computer Science
Publisher
Proceedings of AAAI'99
Published Here
May 3, 2013