What Have We Achieved on Text Summarization?

Huang, Dandan; Cui, Leyang; Yang, Sen; Bao, Guangsheng; Wang, Kun; Xie, Jun; Zhang, Yue

Computer Science > Computation and Language

arXiv:2010.04529 (cs)

[Submitted on 9 Oct 2020]

Title:What Have We Achieved on Text Summarization?

Authors:Dandan Huang, Leyang Cui, Sen Yang, Guangsheng Bao, Kun Wang, Jun Xie, Yue Zhang

View PDF

Abstract:Deep learning has led to significant improvement in text summarization with various methods investigated and improved ROUGE scores reported over the years. However, gaps still exist between summaries produced by automatic summarizers and human professionals. Aiming to gain more understanding of summarization systems with respect to their strengths and limits on a fine-grained syntactic and semantic level, we consult the Multidimensional Quality Metric(MQM) and quantify 8 major sources of errors on 10 representative summarization models manually. Primarily, we find that 1) under similar settings, extractive summarizers are in general better than their abstractive counterparts thanks to strength in faithfulness and factual-consistency; 2) milestone techniques such as copy, coverage and hybrid extractive/abstractive methods do bring specific improvements but also demonstrate limitations; 3) pre-training techniques, and in particular sequence-to-sequence pre-training, are highly effective for improving text summarization, with BART giving the best results.

Comments:	Accepted by EMNLP 2020
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2010.04529 [cs.CL]
	(or arXiv:2010.04529v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2010.04529

Submission history

From: Leyang Cui [view email]
[v1] Fri, 9 Oct 2020 12:39:33 UTC (798 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2020-10

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Dandan Huang
Sen Yang
Kun Wang
Jun Xie
Yue Zhang

export BibTeX citation

Computer Science > Computation and Language

Title:What Have We Achieved on Text Summarization?

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:What Have We Achieved on Text Summarization?

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators