Home › Datasets › modality › Environment
Environment datasets
archive 2025-07-28
147 datasets carry the modality tag "Environment", ordered by the archive's paper count. Page 2 of 4: 48 shown of 147. Facet routes are this site's own (the archive records the tag string, not a page).
The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.
Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets
Environment datasets 49–96 of 147
CARL (Context Adaptive RL)
CARL (context adaptive RL) provides highly configurable contextual extensions to several well-known RL environments.
7 papers · 1 benchmark
Random sampled instances of the Capacitated Vehicle Routing Problem with Time Windows (CVRPTW) for 20, 50 and 100 customer nodes.
7 papers · 0 benchmarks
EvoGym is a large-scale benchmark for co-optimizing the design and control of soft robots.
7 papers · 0 benchmarks
MengeROS is an open-source crowd simulation tool for robot navigation that integrates Menge with ROS.
7 papers · 0 benchmarks
RL Unplugged is suite of benchmarks for offline reinforcement learning.
7 papers · 0 benchmarks
A multimodal agent benchmark on professional data science and engineering.
7 papers · 0 benchmarks
A multivariate spatio-temporal benchmark dataset for meteorological forecasting based on real-time observation data from ground weather stations.
7 papers · 16 benchmarks
AtariARI (Atari Annotated RAM Interface)
The AtariARI (Atari Annotated RAM Interface) is an environment for representation learning.
6 papers · 0 benchmarks
SPACE is a simulator for physical Interactions and causal learning in 3D environments.
6 papers · 0 benchmarks
safe-control-gym is an open-source benchmark suite that extends OpenAI's Gym API with (i) the ability to specify (and query) symbolic models and constraints and (ii) introduce simulated disturbances in the control inputs, measurements, and…
6 papers · 0 benchmarks
FluidLab is a simulation environment with a diverse set of manipulation tasks involving complex fluid dynamics.
5 papers · 0 benchmarks
To collect the 3D Vehicle Tracking Simulation Dataset, a driving simulation is used to obtain accurate 3D bounding box annotations at no cost of human efforts.
4 papers · 0 benchmarks
The DeepMind Alchemy environment is a meta-reinforcement learning benchmark that presents tasks sampled from a task distribution with deep underlying structure.
4 papers · 0 benchmarks
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
4 papers · 0 benchmarks
ClimART (Climate Atmospheric Radiative Transfer)
Numerical simulations of Earth's weather and climate require substantial amounts of computation.
4 papers · 0 benchmarks
LemgoRL is an open-source benchmark tool for traffic signal control designed to train reinforcement learning agents in a highly realistic simulation scenario with the aim to reduce Sim2Real gap.
4 papers · 0 benchmarks
MUAD (Multiple Uncertainties for Autonomous Driving)
The MUAD dataset (Multiple Uncertainties for Autonomous Driving), consisting of 10,413 realistic synthetic images with diverse adverse weather conditions (night, fog, rain, snow), out-of-distribution objects, and annotations for semantic…
4 papers · 0 benchmarks
Phy-Q is a benchmark that requires an agent to reason about physical scenarios and take an action accordingly.
4 papers · 0 benchmarks
The eSports Sensors dataset contains sensor data collected from 10 players in 22 matches in League of Legends.
4 papers · 2 benchmarks
ARCH2S (Dataset, Benchmark for Learning Exterior Architectural Structures from Point Clouds)
Precise segmentation of architectural structures provides detailed information about various building components, enhancing our understanding and interaction with our built environment.
3 papers · 1 benchmark
The 2021 SIGIR workshop on eCommerce is hosting the Coveo Data Challenge for "In-session prediction for purchase intent and recommendations".
3 papers · 1 benchmark
MRPB 1.0 is a mobile robot local planning benchmark.
3 papers · 0 benchmarks
MineRLis an imitation learning dataset with over 60 million frames of recorded human player data.
3 papers · 0 benchmarks
POPGym (Partially Observable Process Gym)
POPGym is designed to benchmark memory in deep reinforcement learning.
3 papers · 0 benchmarks
RoboPianist is a benchmarking suite for high-dimensional control, targeted at testing high spatial and temporal precision, coordination, and planning, all with an underactuated system frequently making-and-breaking contacts.
3 papers · 0 benchmarks
SDN (Situated Dialogue Navigation)
Situated Dialogue Navigation (SDN) is a navigation benchmark of 183 trials with a total of 8415 utterances, around 18.7 hours of control streams, and 2.9 hours of trimmed audio.
3 papers · 0 benchmarks
SILG (Symbolic Interactive Language Grounding)
Symbolic Interactive Language Grounding (SILG) is a multi-environment benchmark which unifies a collection of diverse grounded language learning environments under a common interface.
3 papers · 0 benchmarks
ThreeDWorld Transport Challenge is a visually-guided and physics-driven task-and-motion planning benchmark.
3 papers · 0 benchmarks
This package provides utilities for generation, filtering, solving, visualizing, and processing of mazes for training ML systems.
3 papers · 0 benchmarks
rSoccer is an open-source simulator for the IEEE Very Small Size Soccer and the Small Size League optimized for reinforcement learning experiments.
3 papers · 0 benchmarks
ADORE (A benchmark dataset for machine learning in ecotoxicology)
ADORE is a benchmark dataset for machine learning for ecotixicology, covering acute aquatic toxicity in three relevant taxonomic groups (fish, crustaceans, and algae).
2 papers · 1 benchmark
BASEPROD (The Bardenas Semi-Desert Planetary Rover Dataset)
BASEPROD provides comprehensive rover sensor data collected over a 1.7 km traverse, accompanied by high-resolution 2D and 3D drone maps of the terrain.
2 papers · 0 benchmarks
C2A: Combination to Application Dataset Overview This repository contains the code and information for the paper "UAV-Enhanced Combination to Application: Comprehensive Analysis and Benchmarking of a Human Detection Dataset for Disaster…
2 papers · 1 benchmark
CaFFe (CAlving Fronts and where to Find thEm)
The temporal variability in calving front positions of marine-terminating glaciers permits inference on the frontal ablation.
2 papers · 2 benchmarks
CinemAirSim is an extension of the well-known drone simulator, AirSim, with a cinematic camera as well as extended its API to control all of its parameters in real time, including various filming lenses and common cinematographic…
2 papers · 0 benchmarks
The dataset was created to address the crucial need for effective Extreme Weather Events Detection (EWED), an increasingly urgent task due to the rising frequency of such events driven by global warming.
2 papers · 0 benchmarks
Archive of Global Tropical Cyclone Tracks Tracks from 1980 to May 2019.
2 papers · 0 benchmarks
HeriGraph (Multimodal Machine Learning Datasets on Graphs of Heritage Values and Attributes)
The dataset contains constructed multi-modal features (visual and textual), pseudo-labels (on heritage values and attributes), and graph structures (with temporal, social, and spatial links) constructed using User-Generated Content data…
2 papers · 0 benchmarks
A Simulated Benchmark for multi-modal SLAM Systems Evaluation in Large-scale Dynamic Environments.
2 papers · 0 benchmarks
MMVR (Millimeter-wave Multi-View Radar (MMVR) Dataset)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
2 papers · 0 benchmarks
MSC is a dataset for Macro-Management in StarCraft 2 based on the platfrom SC2LE.
2 papers · 0 benchmarks
Multirotor gym environment for learning control policies for various unmanned aerial vehicles.
2 papers · 0 benchmarks
RL Unplugged is suite of benchmarks for offline reinforcement learning.
2 papers · 0 benchmarks
An environment for RNA design given structure constraints with structures from different datasets to choose from.
2 papers · 0 benchmarks
Large-scale shadows from buildings in a city play an important role in determining the environmental quality of public spaces.
2 papers · 0 benchmarks
SuperCaustics is a simulation tool made in Unreal Engine for generating massive computer vision datasets that include transparent objects.
2 papers · 0 benchmarks
TI1K Dataset (Thumb Index 1000 Hand & Fingertip Detection Dataset)
Thumb Index 1000 (TI1K) is a dataset of 1000 hand images with the hand bounding box, and thumb and index fingertip positions.
2 papers · 0 benchmarks
A procedurally generated jump'n'run game with control over level similarity.
2 papers · 0 benchmarks
Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.