View article

[PDF] from academia.edu

Document summarization using conditional random fields

Authors

Dou Shen, Jian-Tao Sun, Hua Li, Qiang Yang, Zheng Chen

Publication date

2007/1/6

Journal

Proceedings of IJCAI

Volume

Pages

2862-2867

Description

Many methods, including supervised and unsupervised algorithms, have been developed for extractive document summarization. Most supervised methods consider the summarization task as a twoclass classification problem and classify each sentence individually without leveraging the relationship among sentences. The unsupervised methods use heuristic rules to select the most informative sentences into a summary directly, which are hard to generalize. In this paper, we present a Conditional Random Fields (CRF) based framework to keep the merits of the above two kinds of approaches while avoiding their disadvantages. What is more, the proposed framework can take the outcomes of previous methods as features and seamlessly integrate them. The key idea of our approach is to treat the summarization task as a sequence labeling problem. In this view, each document is a sequence of sentences and the summarization procedure labels the sentences by 1 and 0. The label of a sentence depends on the assignment of labels of others. We compared our proposed approach with eight existing methods on an open benchmark data set. The results show that our approach can improve the performance by more than 7.1% and 12.1% over the best supervised baseline and unsupervised baseline respectively in terms of two popular metrics F1 and ROUGE-2. Detailed analysis of the improvement is presented as well.

Total citations

Cited by 463

2007200820092010201120122013201420152016201720182019202020212022202320246 8 23 27 33 28 33 35 34 35 33 34 45 32 16 20 8 5

Scholar articles

Document summarization using conditional random fields.

D Shen, JT Sun, H Li, Q Yang, Z Chen - IJCAI, 2007