Home › Datasets › language › English

English datasets

archive 2025-07-28

3,998 datasets carry the language tag "English", ordered by the archive's paper count. Page 84 of 84: 14 shown of 3,998. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 50 language tags shown of 367, by dataset count; the full filter by modality, task and language is on /datasets

English datasets 3985–3998 of 3,998

0 papers · 0 benchmarks
0 papers · 0 benchmarks
X-Wines (A Wine Dataset for Recommender Systems and Machine Learning)
X-Wines is a consistent wine dataset containing 100,646 instances and 21 million real evaluations carried out by users.
0 papers · 0 benchmarks
Description: The ZakynthosTurtles dataset has been designed to support the development of numerical methods for the recognition and re-identification of individual sea turtles based on their unique scale patterns.
0 papers · 0 benchmarks
ZooScanNet (ZooScanNet: plankton images captured with the ZooScan)
Plankton was sampled with various nets, from bottom or 500m depth to the surface, in many oceans of the world.
0 papers · 0 benchmarks
eAppleScab (Apple Scab in the Early Stage of Development)
The study showed that the apple scab can be detected in the high-resolution RGB images in an early stage of its development.
0 papers · 0 benchmarks
fNIRS2MW (The Tufts fNIRS to Mental Workload Dataset)
The Tufts fNIRS to Mental Workload (fNIRS2MW) open-access dataset is a new dataset for building machine learning classifiers that can consume a short window (30 seconds) of multivariate fNIRS recordings and predict the mental workload…
0 papers · 0 benchmarks
fish (fishway)
The data was captured from an overhead perspective, showcasing the swimming behavior of fish in a simulated flowing water channel.
0 papers · 0 benchmarks
maadaa-FaEco Dataset (maadaa.ai Fashion & e-Commerce Open Dataset)
The dataset is organized into 24 typical scenarios, showcasing the richness of real-world environments, conditions, and objects.
0 papers · 0 benchmarks
test-dataset
0 papers · 0 benchmarks
test10
0 papers · 0 benchmarks
tomato detection (A dataset of tomato fruits images for object detection in the complex lighting environment of plant factories)
Plant factories are an advanced form of facility agriculture that enable efficient plant cultivation through controllable environmental conditions, making them highly suitable for the automation and intelligent application of machinery.
0 papers · 0 benchmarks
tomato fruits detection (A dataset of tomato fruits images for object detection in the complex lighting environment of plant factories)
Plant factories are an advanced form of facility agriculture that enable efficient plant cultivation through controllable environmental conditions, making them highly suitable for the automation and intelligent application of machinery.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.