Home › Datasets › language › English

English datasets

archive 2025-07-28

3,998 datasets carry the language tag "English", ordered by the archive's paper count. Page 43 of 84: 48 shown of 3,998. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 language tags shown of 367, by dataset count; the full filter by modality, task and language is on /datasets

English datasets 2017–2064 of 3,998

The QTUNA dataset is the result of a series of elicitation experiments in which human speakers were asked to perform a linguistic task that invites the use of quantified expressions in order to inform possible Natural Language Generation…
2 papers · 0 benchmarks
QUITE (Quantifying Uncertainty in natural language Text)
QUITE (Quantifying Uncertainty in natural language Text) is an entirely new benchmark that allows for assessing the capabilities of neural language model-based systems w.r.t.
2 papers · 0 benchmarks
Introduction Generalized quantifiers (e.g., few, most) are used to indicate the proportions predicates are satisfied.
2 papers · 0 benchmarks
R2VQ (Recipe-to-Video Questions)
R2VQ is a dataset designed for testing competence-based comprehension of machines over a multimodal recipe collection, which contains text-video aligned recipes.
2 papers · 0 benchmarks
A human-revised dataset for seven languages that allows for the evaluation of multilingual RE systems.
2 papers · 0 benchmarks
RGBD1K (A Large-scale Dataset and Benchmark for RGB-D Object Tracking)
RGBD1K is a benchmark for RGB-D Object Tracking which contains 1050 sequences with about 2.5M frames in total.
2 papers · 0 benchmarks
RIR dataset (Planar Room Impulse Response Dataset - ACT, DTU Electro (b. 355 r. 008))
Dataset of Room Impulse Responses measured at the Acoustic Technology group facilities, DTU Electro.
2 papers · 0 benchmarks
RISEdb (Robust Indoor Localization in Complex Scenarios (RISE) database)
The RISE (Robust Indoor Localization in Complex Scenarios) dataset is meant to train and evaluate visual indoor place recognizers.
2 papers · 0 benchmarks
RLU (RL Unplugged)
RL Unplugged is suite of benchmarks for offline reinforcement learning.
2 papers · 0 benchmarks
The first benchmark comprising 473 prompts designed to assess the ability of LLMs to resist malicious code generation.
2 papers · 0 benchmarks
ROPE (Recognition-based Object Probing Evaluation)
We introduce Recognition-based Object Probing Evaluation (ROPE), an automated evaluation protocol that considers the distribution of object classes within a single image during testing and uses visual referring prompts to eliminate…
2 papers · 0 benchmarks
RTC (Reddit Time Corpus)
RTC is a benchmark corpus of social media comments sampled over three years.
2 papers · 0 benchmarks
RUFF is a large-scale dataset to measure pronoun fidelity in English.
2 papers · 0 benchmarks
RUGD (RUGD: Robot Unstructured Ground Driving)
A Video Dataset for Visual Perception and Autonomous Navigation in Unstructured Environments.
2 papers · 1 benchmark
The RaDelft dataset is a novel, large-scale, real-life, and multi-sensor dataset that has been recorded using a demonstrator vehicle in different locations in the city of Delft.
2 papers · 0 benchmarks
RaidaR (RaidaR: A Rich Annotated Image Dataset of Rainy Street Scenes)
RaidaR is a rich annotated image dataset of rainy street scenes.
2 papers · 0 benchmarks
RaindropClarity (A Dual-Focused Dataset for Day and Night Raindrop Removal)
Existing raindrop removal datasets have two shortcomings.
2 papers · 0 benchmarks
A dataset consisting of recipient 46 users and, 26180 tweets.
2 papers · 0 benchmarks
Dataset used in the publication of Rapid Design of Top-Performing Metal-Organic Frameworks with Qualitative Representations of Building Blocks.
2 papers · 0 benchmarks
Rare Diseases Mentions in MIMIC-III (Rare disease mention annotations from a sample of MIMIC-III clinical notes)
Data annotation The 1,073 full rare disease mention annotations (from 312 MIMIC-III discharge summaries) are in fullsetRDannMIMICIIIdisch.csv.
2 papers · 1 benchmark
ReactionGIF is an affective dataset of 30K tweets which can be used for tasks like induced sentiment prediction and multilabel classification of induced emotions.
2 papers · 0 benchmarks
Articles originating from subreddits with explicitly stated ideologies are categorized into three groups: 72,488 articles in the Liberal class, 79,573 articles in the Conservative class, and 225,083 articles in the Restricted class.
2 papers · 1 benchmark
Rent3D++ is an extension of the Rent3D floorplans + photos dataset.
2 papers · 1 benchmark
Transaction fee mechanism (TFM) is an essential component of a blockchain protocol.
2 papers · 0 benchmarks
Noiseless reverberant dataset using the public WSJ0 corpus and simulated room impulse responses using the PyRoomAcoustics library.
2 papers · 0 benchmarks
RoFT-chatgpt is a variation of RoFT dataset, where the same human prompts are continued with the gpt-3.5-turbo model.
2 papers · 1 benchmark
RoMQA is a benchmark for robust, multi-evidence, and multi-answer question answering (QA).
2 papers · 0 benchmarks
RyanSpeech is a speech corpus for research on automated text-to-speech (TTS) systems.
2 papers · 0 benchmarks
Our proposed Synthetic-to-Real benchmark for more practical visual DA (termed S2RDA) includes two challenging transfer tasks of S2RDA-49 and S2RDA-MS-39.
2 papers · 0 benchmarks
SA-Det-100k is a large-scale class-agnostic object detection dataset for Research Purposes only.
2 papers · 1 benchmark
SEED-VIG (SJTU Emotion EEG Dataset)
The SEED-VIG dataset is composed of four parts.
2 papers · 0 benchmarks
SEPE 8K dataset is made of 40 different 8K (8192 x 4320) video sequences and 40 variant 8K (8192 x 5464) images.
2 papers · 1 benchmark
SLAM2REF (ConSLAM BIM and GT Poses)
This dataset comprehends the 3D building information model (in IFC and Revit formats), manually elaborated based on the terrestrial laser scanner of the sequence 2 of ConSLAM, and the refined ground truth (GT) poses (in TUM format) of…
2 papers · 0 benchmarks
SQL-Eval is an open-source PostgreSQL evaluation dataset released by Defog, constructed based on Spider.
2 papers · 1 benchmark
SSL4EO-S12 is a large-scale, global, multimodal, and multi-seasonal corpus of satellite imagery from the ESA Sentinel-1 & -2 satellite missions.
2 papers · 0 benchmarks
SVBench (Streaming Video Understanding Benchmark)
Dataset Card for SVBench This dataset card aims to provide a comprehensive overview of the SVBench dataset, including its purpose, structure, and sources.
2 papers · 0 benchmarks
Sakuga-42M is a large-scale hand-drawn cartoon video dataset for academic research purposes, it comprises 42 million cartoon keyframes covering various artistic styles, regions, and years, with comprehensive semantic annotations including…
2 papers · 0 benchmarks
SciGen is a challenge dataset for the task of reasoning-aware data-to-text generation consisting of tables from scientific articles and their corresponding descriptions.
2 papers · 0 benchmarks
The SegmentedTables dataset is a collection of almost 2,000 tables extracted from 352 machine learning papers.
2 papers · 0 benchmarks
The SheetCopilot dataset contains 28 evaluation workbooks and 221 spreadsheet manipulation tasks that are applied to these workbooks.
2 papers · 1 benchmark
Recent applications of LLMs in Machine Reading Comprehension (MRC) systems have shown impressive results, but the use of shortcuts, mechanisms triggered by features spuriously correlated to the true label, has emerged as a potential threat…
2 papers · 0 benchmarks
LLMs' lateral thinking capabilities remain under-explored and challenging to measure due to the complexity of assessing creative thought processes and the scarcity of relevant data.
2 papers · 0 benchmarks
Social media attributions of YouTube comments (Social media attributions dataset of YouTube comments in the context of water crisis)
Data set constructed from YouTube comments (72,098 comments posted by 43,859 users on 623 relevant videos to the crisis)
2 papers · 1 benchmark
https://huggingface.co/papers/2502.20730
2 papers · 0 benchmarks
Sparrow (Sparrow-V0: A Reinforcement Learning Friendly Simulator for Mobile Robot)
Sparrow-V0: A Reinforcement Learning Friendly Simulator for Mobile Robot Features: Vectorizable (Enable fast data collection; Single environment is also supported) Domain Randomization (control interval, control delay, maximum velocity,…
2 papers · 0 benchmarks
Speech Accent Archive (The Speech Accent Archive)
The speech accent archive uniformly presents a large set of speech samples from a variety of language backgrounds.
2 papers · 1 benchmark
This resource is designed to allow for research into Natural Language Generation.
2 papers · 0 benchmarks
Electrophysiological data from implanted electrodes in the human brain are rare, and therefore scientific access to it has remained somewhat exclusive.
2 papers · 1 benchmark

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.