Home › Datasets
Datasets
archive 2025-07-28
12,172 datasets listed, ordered by the archive's paper count. Page 242 of 254: 48 shown of 12,172.
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter a dataset can carry several tags, so counts overlap
Modality 39
Task 500 shown of 3,717, by dataset count
Language 367
All datasets 11569–11616 of 12,172
Freiburg Across Seasons captures long-term perceptual changes across a span of 3 years.
0 papers · 0 benchmarks
Freiburg Block Tasks is a dataset for robot skill learning.
0 papers · 0 benchmarks
The Freiburg Campus 3D Scan dataset consists of 3D area maps from the Freiburg campus that were scanned with 3D lasers.
0 papers · 0 benchmarks
Freiburg Lighting Adaptable Map Tracking is a dataset for camera trajectory estimation.
0 papers · 0 benchmarks
The Freiburg Poking dataset is a dataset for learning intuitive physics from physical interaction.
0 papers · 0 benchmarks
The Freiburg RGB-D People dataset contains 3000+ RGB-D frames acquired in a university hall from three vertically mounted Kinect sensors.
0 papers · 0 benchmarks
Description: Replication data and code for Aref, S., and Neal, Z.P., "Identifying hidden coalitions in the US House of Representatives by optimally partitioning signed networks based on generalized balance" (2021) Scientific Reports.
0 papers · 0 benchmarks
GDXray+ is a collection of more than 21.100 X-ray images for the development, testing, and evaluation of image analysis and computer vision algorithms.
0 papers · 0 benchmarks
GNMC (Gracenote Multi-Crop Dataset)
We present the Gracenote Multi-Crop (GNMC) dataset, to further research in algorithms for aesthetic image cropping.
0 papers · 0 benchmarks
This record contains the saddle search output logs for Sella and EON (dimer, with and without GPR acceleration).
0 papers · 0 benchmarks
Gap Pattern Detection (Gap Pattern (Gap Up and Gap Down) Detection in Candlestick Trading Charts for Technical Analysis)
1.
0 papers · 0 benchmarks
GenAI-Bench benchmark consists of 1,600 challenging real-world text prompts sourced from professional designers.
0 papers · 0 benchmarks
We've made available several genome-wide datasets, which can be used for training microRNA (miRNA) classifiers.
0 papers · 0 benchmarks
A large database of geotagged face images.
0 papers · 0 benchmarks
German affixoids are a type of morpheme in between affixes and free stems.
0 papers · 0 benchmarks
Corpus and annotations for the CL-Aff Shared Task - Get it #OffMyChest - from Nanyang Technological University Singapore.
0 papers · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
0 papers · 0 benchmarks
Google Research Footbal is a new reinforcement learning environment where agents are trained to play football in an advanced, physics-based 3D simulator.
0 papers · 0 benchmarks
The Grand central station dataset includes a video with 50,010 frames which is used for Scene Understanding and Crowd Analysis.
0 papers · 0 benchmarks
KOKLU Murat (a), UNLERSEN M.
0 papers · 0 benchmarks
Grévy’s Zebra is an animal pose estimation dataset for zebras.
0 papers · 0 benchmarks
HEADSET (HEADSET: Human Emotion Awareness under Partial Occlusions Multimodal DataSET)
The volumetric representation of human interactions is one of the fundamental domains in the development of immersive media productions and telecommunication applications.
0 papers · 0 benchmarks
Sample data in the numpy array format (.npy) at the link https://zenodo.org/record/7189381#.Y0a2UHZBxD9.
0 papers · 0 benchmarks
HWR200 (New open access dataset of handwritten texts images in Russian)
New open access dataset of handwritten texts images in Russian
0 papers · 0 benchmarks
The Hacker News dataset provides a valuable glimpse into the tech industry's landscape.
0 papers · 0 benchmarks
This dataset comprises the raw data of four three-dimensional seismic surveys acquired in Canadian mining camps.
0 papers · 0 benchmarks
The medaka (Oryzias latipes) and the zebrafish (Danio rerio) are used as a model organism for a variety of subjects in biomedical research.
0 papers · 0 benchmarks
Heel Bone X-Ray Dataset consists of 3,956 X-ray images of the foot, primarily focused on detecting and classifying heel bone diseases.
0 papers · 0 benchmarks
This dataset is an extremely challenging set of over 5000+ original Hindi text images captured and crowdsourced from over 700+ urban and rural areas, where each image is manually reviewed and verified by computer vision professionals at…
0 papers · 0 benchmarks
We have constructed our dataset by five fields available on the website that were found convenient for the study of student expectations and experience.
0 papers · 0 benchmarks
The Hong Kong Cantonese Corpus was collected from transcribed conversations that were recorded between March 1997 and August 1998.
0 papers · 0 benchmarks
HouseCat6D (A Large-Scale Multi-Modal Category Level 6D Object Perception Dataset with Household Objects in Realistic Scenarios)
Estimating 6D object poses is a major challenge in 3D computer vision.
0 papers · 0 benchmarks
HuSc3D (Human Sculpture dataset for 3D object reconstruction)
HuSc3D is a novel dataset specifically designed for rigorous benchmarking of 3D reconstruction models under realistic acquisition challenges.
0 papers · 0 benchmarks
Huawei University Challenge Competition 2021 Data Science for Indoor positioning 2.2 Full Mall Graph Clustering Train The sample training data for this problem is a set of 106981 fingerprints (task2trainfingerprints.json) and some edges…
0 papers · 0 benchmarks
The dataset consists of images of Human palms captured using a mobile phone.
0 papers · 0 benchmarks
This dataset consists of images of wrist (with different kind of bands on it).
0 papers · 0 benchmarks
This dataset consists of 600+ items of faces with different emotions and mixed races that are ready to use for optimizing the accuracy of computer vision models.
0 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.