{"url":"/dataset/value","name":"VALUE","full_name":"Video-And-Language Understanding Evaluation","description_markdown":"**VALUE** is a Video-And-Language Understanding Evaluation benchmark to test models that are generalizable to diverse tasks, domains, and datasets. It is an assemblage of 11 VidL (video-and-language) datasets over 3 popular tasks: (i) text-to-video retrieval; (ii) video question answering; and (iii) video captioning. VALUE benchmark aims to cover a broad range of video genres, video lengths, data volumes, and task difficulty levels. Rather than focusing on single-channel videos with visual information only, VALUE promotes models that leverage information from both video frames and their associated subtitles, as well as models that share knowledge across multiple tasks. \r\n\r\nThe datasets used for the VALUE benchmark are: [TVQA](tvqa), [TVR](tvr), [TVC](tvc), [How2R](how2r), [How2QA](how2qa), [VIOLIN](violin), [VLEP](vlep), [YouCook2](youcook2) (YC2C, YC2R), [VATEX](vatex)","description_withheld":null,"homepage":"https://value-leaderboard.github.io/","introduced_date":"2021-06-08","introduced_date_note":null,"introduced_by":{"paper":"/paper/value-a-multi-task-benchmark-for-video-and","title":"VALUE: A Multi-Task Benchmark for Video-and-Language Understanding Evaluation","first_author":"Linjie Li","url":null},"license":{"name":"Multiple licenses","url":null},"modalities":[{"name":"Videos","url":"/datasets/modality/videos"},{"name":"Texts","url":"/datasets/modality/texts"}],"tasks":[{"name":"Video Question Answering","url":"/task/video-question-answering","datasets_with_task":"/datasets/task/video-question-answering"},{"name":"Video Captioning","url":"/task/video-captioning","datasets_with_task":"/datasets/task/video-captioning"},{"name":"Text-to-video search","url":"/task/text-to-video-search","datasets_with_task":"/datasets/task/text-to-video-search"}],"languages":[{"name":"English","url":"/datasets/language/english"},{"name":"Chinese","url":"/datasets/language/chinese"}],"variants":["VALUE"],"data_loaders":[],"num_papers_in_archive":30,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}