{"url":"/dataset/ami-meeting-corpus","name":"AMI Meeting Corpus","full_name":null,"description_markdown":"The **AMI Meeting Corpus** is a **multi-modal data set** comprising **100 hours of meeting recordings**. It has been meticulously curated for research purposes and includes various modes of data capture. Let me provide you with more details:\r\n\r\n1. **Purpose and Context**:\r\n   - The corpus was created in the context of a project that aims to develop **meeting browsing technology**.\r\n   - Eventually, it will be **publicly released** for use by researchers and practitioners.\r\n\r\n2. **Data Composition**:\r\n   - **100 hours** of recorded meetings are included.\r\n   - The data is collected from **various domains** and scenarios.\r\n   - Around **two-thirds** of the data involves participants playing different roles in a **design team**, taking a design project from kick-off to completion over the course of a day.\r\n   - The remaining data consists of **naturally occurring meetings**.\r\n\r\n3. **Modalities and Annotations**:\r\n   - The corpus includes **synchronized recording devices** such as close-talking and far-field microphones, individual and room-view video cameras, projection, a whiteboard, and individual pens.\r\n   - Annotations cover various phenomena, including **orthographic transcription**, **dialog acts**, and **head movement**.\r\n\r\n4. **Research Applications**:\r\n   - Although initially designed for meeting browsing technology, the AMI Meeting Corpus is **useful for a wide range of research areas**.\r\n   - Researchers engaged in **video processing** can access higher resolution videos.\r\n\r\nSource: Conversation with Bing, 3/16/2024\r\n(1) AMI Corpus - University of Edinburgh. https://groups.inf.ed.ac.uk/ami/corpus/.\r\n(2) The AMI Meeting Corpus: A Pre-announcement. https://www.research.ed.ac.uk/en/publications/the-ami-meeting-corpus-a-pre-announcement.\r\n(3) The AMI Meeting Corpus: A Pre-announcement | SpringerLink. https://link.springer.com/chapter/10.1007/11677482_3.\r\n(4) The AMI Meeting Corpus: A Pre-announcement — University of Twente .... https://research.utwente.nl/en/publications/the-ami-meeting-corpus-a-pre-announcement.","description_withheld":null,"homepage":"https://groups.inf.ed.ac.uk/ami/corpus","introduced_date":null,"introduced_date_note":null,"introduced_by":null,"license":null,"modalities":[{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Meeting Summarization","url":"/task/meeting-summarization","datasets_with_task":"/datasets/task/meeting-summarization"},{"name":"Abstractive Dialogue Summarization","url":"/task/abstractive-dialogue-summarization","datasets_with_task":"/datasets/task/abstractive-dialogue-summarization"}],"languages":[{"name":"English","url":"/datasets/language/english"}],"variants":["AMI Meeting Corpus"],"data_loaders":[],"num_papers_in_archive":6,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[{"leaderboard":"/sota/meeting-summarization-on-ami-meeting-corpus","task":"Meeting Summarization","dataset_variant":"AMI Meeting Corpus","rows":1,"metrics":["ROUGE-1 F1"],"first_row_in_archive_order":{"model":"UNS","paper":"/paper/unsupervised-abstractive-meeting","metrics":{"ROUGE-1 F1":"37.53"},"code_links":[{"title":"xcfcode/Summarization-Papers","url":"https://github.com/xcfcode/Summarization-Papers"},{"title":"Tixierae/gow_tools","url":"https://github.com/Tixierae/gow_tools"},{"title":"bearblog/CoreRank","url":"https://github.com/bearblog/CoreRank"},{"title":"dascim/acl2018_abssumm","url":"https://bitbucket.org/dascim/acl2018_abssumm"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"}],"papers_with_a_benchmark_row":[{"paper":"/paper/unsupervised-abstractive-meeting","title":"Unsupervised Abstractive Meeting Summarization with Multi-Sentence Compression and Budgeted Submodular Maximization","date":"2018-05-14","rows_on_this_dataset":1,"code_links":4,"syntology":null}],"syntology_totals":{"read_at":"2026-09-25T09:33:49+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}