Home › Datasets › modality › Interactive
Interactive datasets
archive 2025-07-28
29 datasets carry the modality tag "Interactive", ordered by the archive's paper count. Page 1 of 1: 29 shown of 29. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Interactive datasets 1–29 of 29
The SUN RGBD dataset contains 10335 real RGB-D images of room scenes.
477 papers · 11 benchmarks
ZINC is a free database of commercially-available compounds for virtual screening.
251 papers · 5 benchmarks
R2R is a dataset for visually-grounded natural language navigation in real buildings.
174 papers · 2 benchmarks
NCLT (North Campus Long-Term Vision and LiDAR)
The NCLT dataset is a large scale, long-term autonomy dataset for robotics research collected on the University of Michigan’s North Campus.
124 papers · 0 benchmarks
The MAESTRO dataset contains over 200 hours of paired audio and MIDI recordings from ten years of International Piano-e-Competition.
118 papers · 1 benchmark
CoMA contains 17,794 meshes of the human face in various expressions Source: DEMEA: Deep Mesh Autoencoders for Non-Rigidly Deforming Objects Image Source: https://coma.is.tue.mpg.de/
81 papers · 1 benchmark
Jericho is a learning environment for man-made Interactive Fiction (IF) games.
54 papers · 0 benchmarks
The Kumar dataset contains 30 1,000×1,000 image tiles from seven organs (6 breast, 6 liver, 6 kidney, 6 prostate, 2 bladder, 2 colon and 2 stomach) of The Cancer Genome Atlas (TCGA) database acquired at 40× magnification.
17 papers · 1 benchmark
The NYU Hand pose dataset contains 8252 test-set and 72757 training-set frames of captured RGBD data with ground-truth hand-pose information.
16 papers · 1 benchmark
The Watch-n-Patch dataset was created with the focus on modeling human activities, comprising multiple actions in a completely unsupervised setting.
13 papers · 0 benchmarks
Synbols is a dataset generator designed for probing the behavior of learning algorithms.
11 papers · 0 benchmarks
DCASE2018 Task 4 is a dataset for large-scale weakly labeled semi-supervised sound event detection in domestic environments.
10 papers · 0 benchmarks
ViP-Bench (Making Large Multimodal Models Understand Arbitrary Visual Prompts)
ViP-Bench is a comprehensive benchmark designed to assess the capability of multimodal models in understanding visual prompts across multiple dimensions.
10 papers · 1 benchmark
A multimodal agent benchmark on professional data science and engineering.
7 papers · 0 benchmarks
Avalon is a benchmark for generalization in Reinforcement Learning (RL).
5 papers · 0 benchmarks
7,672 human written natural language navigation instructions for routes in OpenStreetMap with a focus on visual landmarks.
5 papers · 2 benchmarks
The Situated Corpus Of Understanding Transactions (SCOUT) is a multi-modal collection of human-robot dialogue in the task domain of collaborative exploration.
4 papers · 0 benchmarks
MiniWob++ is a suite of web-browser based tasks introduced in Liu et al.
2 papers · 0 benchmarks
SuperCaustics is a simulation tool made in Unreal Engine for generating massive computer vision datasets that include transparent objects.
2 papers · 0 benchmarks
The bipedal skills benchmark is a suite of reinforcement learning environments implemented for the MuJoCo physics simulator.
2 papers · 0 benchmarks
In this dataset two robots, Baxter and UR5, perform 8 behaviors (look, grasp, pick, hold, shake, lower, drop, and push) on 95 objects that vary by 5 color (blue, green, red, white, and yellow), 6 contents (wooden button, plastic dices,…
1 paper · 0 benchmarks
In this dataset an uppertorso humanoid robot with 7-DOF arm explored 100 different objects belonging to 20 different categories using 10 behaviors: Look, Crush, Grasp, Hold, Lift, Drop, Poke, Push, Shake and Tap.
1 paper · 0 benchmarks
A large-scale, first-of-its-kind database aimed at generating a better understanding of the way children interact with mobile devices during their development process.
1 paper · 0 benchmarks
The Daimler Monocular Pedestrian Detection dataset is a dataset for pedestrian detection in urban environments.
1 paper · 0 benchmarks
We conducted a large crowdsourcing study of click patterns in an interactive segmentation scenario and collected 475K real-user clicks.
1 paper · 0 benchmarks
We use Something-Something v2 dataset to obtain the generation prompts and ground truth masks from real action videos.
1 paper · 0 benchmarks
The Mafia Dataset was created to model the behavior of deceptive actors in the context of the Mafia game, as described in the paper “Putting the Con in Context: Identifying Deceptive Actors in the Game of Mafia”.
1 paper · 0 benchmarks
In this dataset UR5 robot used 6 tools: metal-scissor, metal-whisk, plastic-knife, plastic-spoon, wooden-chopstick, and wooden-fork to perform 6 behaviors: look, stirring-slow, stirring-fast, stirring-twist, whisk, and poke.
1 paper · 0 benchmarks
The datasets includes curves drawn on 3D surfaces (triangle meshes) in Virtual Reality.
1 paper · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.