1999 Presentations (Communicative Events)
Towards multidocument summarization by reformulation: progress and prospects
By synthesizing information common to retrieved documents, multi-document summarization can help users of information retrieval systems to find relevant documents with a minimal amount of reading. We are developing a multidocument summarization system to automatically generate a concise summary by identifying and synthesizing similarities across a set of related documents. Our approach is unique in its integration of machine learning and statistical techniques to identify similar paragraphs, intersection of similar phrases within paragraphs, and language generation to reformulate the wording of the summary. Our evaluation of system components shows that learning over multiple extracted linguistic features is more effective than information retrieval approaches at identifying similar text units for summarization and that it is possible to generate a fluent summary that conveys similarities among documents even when full semantic interpretations of the input text are not available.
Subjects
Files
- mckeown_al_99.pdf application/pdf 85.2 KB Download File
More About This Work
- Academic Units
- Computer Science
- Publisher
- Proceedings of AAAI'99
- Published Here
- May 3, 2013