{"url":"/sota/document-summarization-on-arxiv-summarization","task":{"name":"Document Summarization","url":"/task/document-summarization","note":null},"dataset":{"name":"arXiv Summarization Dataset","url":"/dataset/arxiv-summarization-dataset"},"category":"Natural Language Processing","categories":["Natural Language Processing"],"category_note":null,"description":"Automatic **Document Summarization** is the task of rewriting a document into its shorter form while still retaining its important content. The most popular two paradigms are extractive approaches and abstractive approaches. Extractive approaches generate summaries by extracting parts of the original document (usually sentences), while abstractive methods may generate new words or phrases which are not in the original document.\n\n\n<span class=\"description-source\">Source: [HIBERT: Document Level Pre-training of Hierarchical Bidirectional Transformers for Document Summarization ](https://arxiv.org/abs/1905.06566)</span>","description_from":"task","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","rank":"the archive's row order at snapshot; not re-ranked","rows_end_at":"2025-07-28","rows_withheld_as_spam":0,"metric_values":"the archive's strings, untouched"},"metrics":["Rouge-2"],"metric_direction":{"note":"inferred from the metric name only (the archive records no direction); null = not inferred, chart draws points only","by_metric":{"Rouge-2":"higher"}},"counts":{"rows":1,"rows_with_code":0,"rows_with_paper_page":1,"rows_dated":1,"rows_using_additional_data":0},"rows":[{"rank_in_archive_order":1,"model":"DeepPyramidion","metrics":{"Rouge-2":"19.99"},"uses_additional_data":false,"paper_date":"2021-11-16","paper":"/paper/sparsifying-transformer-models-with-trainable","paper_url":"https://openreview.net/forum?id=c3wXWy6xe3O","paper_title":"Sparsifying Transformer Models with Trainable Representation Pooling","code":null,"n_code_links":0,"syntology":null}],"since_archive":{"present":false,"note":"No Syntology-extracted rows are published in this build."},"syntology":{"read_at":"2026-09-24T18:15:14+00:00","claim":"Per row: N of M harvested code samples from that row's paper executed on a synthesized fixture; the other M-N are unverified. Not a reproduction of the row's number; not a correctness claim. n_pointer_only_licence counts samples the site points at rather than redistributes (a licence axis, independent of ran/unverified).","rows_with_graph_line":0,"rows_with_any_sample_ran":0,"distinct_papers_with_graph_line":0,"distinct_papers_with_any_sample_ran":0,"samples_over_distinct_papers":{"n_ran":0,"n_unverified":0,"n_samples":0,"n_pointer_only_licence":0,"note":"each paper (arXiv id) counted once, however many rows it is behind; this is the page-level figure"},"samples_row_weighted":{"n_ran":0,"n_unverified":0,"n_samples":0,"n_pointer_only_licence":0,"note":"row-weighted: a paper behind several rows is counted once per row; inflated relative to samples_over_distinct_papers by design, kept for readers summing the per-row syntology blocks"}}}