Papers › Controlling the Amount of Verbatim Copying in Abstractive Summarization

Controlling the Amount of Verbatim Copying in Abstractive Summarization

23 Nov 2019arXiv:1911.10390archive 2025-07-28

Kaiqiang Song, Bingqing Wang, Zhe Feng, Liu Ren, Fei Liu

An abstract must not change the meaning of the original text. A single most effective way to achieve that is to increase the amount of copying while still allowing for text abstraction. Human editors can usually exercise control over copying, resulting in summaries that are more extractive than abstractive, or vice versa. However, it remains poorly understood whether modern neural abstractive summarizers can provide the same flexibility, i.e., learning from single reference summaries to generate multiple summary hypotheses with varying degrees of copying. In this paper, we present a neural summarization model that, by learning from single human abstracts, can produce a broad spectrum of summaries ranging from purely extractive to highly generative ones. We frame the task of summarization as language modeling and exploit alternative mechanisms to generate summary hypotheses. Our method allows for control over copying during both training and decoding stages of a neural summarization model. Through extensive experiments we illustrate the significance of our proposed method on controlling the amount of verbatim copying and achieve competitive results over strong baselines. Our analysis further reveals interesting and unobvious facts.

PaperPDFCode

In Syntology View this paper on Syntology: its repositories, every harvested function with whether it ran, its licence and the call to fetch it.

Open this paper in Syntology's Atlas, the map of the papers in Syntology's graph and their citations.

Code

ucfnlp/control-over-copying officialmentioned in papermentioned on GitHubpytorch report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

Abstractive Text SummarizationLanguage ModelingLanguage ModellingText Summarization

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Text Summarization GigaWord ControlCopying + BPNorm ROUGE-1 39.19 #14 of 41 Archive leaderboard report
Text Summarization GigaWord ControlCopying + BPNorm ROUGE-2 20.38 #14 of 41 Archive leaderboard report
Text Summarization GigaWord ControlCopying + BPNorm ROUGE-L 36.69 #14 of 41 Archive leaderboard report
Text Summarization GigaWord ControlCopying + SBWR ROUGE-1 39.08 #17 of 41 Archive leaderboard report
Text Summarization GigaWord ControlCopying + SBWR ROUGE-2 20.47 #17 of 41 Archive leaderboard report
Text Summarization GigaWord ControlCopying + SBWR ROUGE-L 36.69 #17 of 41 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections