Home › Datasets › modality › Environment

Environment datasets

archive 2025-07-28

147 datasets carry the modality tag "Environment", ordered by the archive's paper count. Page 3 of 4: 48 shown of 147. Facet routes are this site's own (the archive records the tag string, not a page).

The archive holds 12,214 dataset rows; 12,172 are listed. 6 are withheld from every listing and count here as vandalised before snapshot (6 with contact-centre spam in the title, 0 with a spam description on a row that has no homepage, no paper and no papers counted; none with more than 1 paper, 0 with a benchmark), listed in withheld.json; 1 listed row carries a vandalised description, withheld on its page. This gate never withholds a row with a homepage or a paper that resolves, and a clean description; the content rules below withhold a row whose name is spam whatever else it carries. The gate is a phrase list: these are the rows it caught, not a claim that the rest is clean. Before that gate, the site's content rules withhold 36 more rows (invite-code, gambling, travel-booking, contact-centre and similar spam in the name or on a row with nothing real behind it); they have no page and are listed in withheld.json.

Filter 39 modality tags shown of 39, by dataset count; the full filter by modality, task and language is on /datasets

Environment datasets 97–144 of 147

TeachMyAgent (TA) is a benchmark for Automatic Curriculum Learning (ACL) algorithms leveraging procedural task generation.
2 papers · 0 benchmarks
The 2048 game task involves training an agent to achieve high scores in the game 2048 (Wikipedia))
2 papers · 1 benchmark
Wastewater catchment area data are essential for wastewater treatment capacity planning and have recently become critical for operationalising wastewater-based epidemiology (WBE) for COVID-19.
2 papers · 0 benchmarks
bipedal-skills (Bipedal Skills Benchmark for Reinforcement Learning)
The bipedal skills benchmark is a suite of reinforcement learning environments implemented for the MuJoCo physics simulator.
2 papers · 0 benchmarks
Unsustainable fishing practices worldwide pose a major threat to marine resources and ecosystems.
2 papers · 1 benchmark
TDW is a 3D virtual world simulation platform, utilizing state-of-the-art video game engine technology.
1 paper · 0 benchmarks
This dataset is meant to be used to develop models for next-day fire hazard forecasting in Greece.
1 paper · 0 benchmarks
BIRDeep (BIRDeep_AudioAnnotations)
The BIRDeep Audio Annotations dataset is a collection of bird vocalizations from Doñana National Park, Spain.
1 paper · 0 benchmarks
A dataset for flying honeybee detection introduced in "A Method for Detection of Small Moving Objects in UAV Videos".
1 paper · 1 benchmark
BrazilDAM is a multi sensor and multitemporal dataset that consists of multispectral images of ore tailings dams throughout Brazil.
1 paper · 0 benchmarks
BurnMD (A Fire Projection and Mitigation Modeling Dataset)
A dataset composed of 308 medium sized fires from the years 2018-2021, complete with both time series airborne based inference and ground operational estimation of fire extent, and operational mitigation data such as control line…
1 paper · 0 benchmarks
This dataset contains the bus trajectory dataset collected by 6 volunteers who were asked to travel across the sub-urban city of Durgapur, India, on intra-city buses (route name: 54 Feet).
1 paper · 0 benchmarks
CARLE (Cellular Automata Reinforcement Learning Environment)
CARLE is a life-like cellular automata simulator and reinforcement learning environment.
1 paper · 0 benchmarks
Given the difficulty to handle planetary data we provide downloadable files in PNG format from the missions Chang'E-3 and Chang'E-4.
1 paper · 0 benchmarks
ColosseumRL is a framework for research in reinforcement learning in n-player games.
1 paper · 0 benchmarks
The 50-ha plot at Barro Colorado Island was initially demarcated and fully censused in 1982, and has been fully censused 7 times since, every 5 years from 1985 through 2015.
1 paper · 0 benchmarks
DARai (Daily Activity Recordings for AI and ML applications)
Daily Activity Recordings for Artificial Intelligence (DARai, pronounced "Dahr-ree") is a multimodal, hierarchically annotated dataset constructed to understand human activities in real-world settings.
1 paper · 0 benchmarks
[comment]:<> (Data for the paper "Deciphering Environmental Air Pollution with Large Scale City Data") Main Dataset citypollutiondata.csv Relevant Columns: Date: Date of the sample City: City of the sample Xmedian: Median value of the…
1 paper · 0 benchmarks
We release both the processed data and evaluation results from our own experiments, and the underlying raw data that can be used for future experiments and schemes in the domain of Zero-Interaction Security.
1 paper · 0 benchmarks
This dataset contains recordings of 32 sound producing insect species with a total 335 files and a length of 57 minutes.
1 paper · 0 benchmarks
Deep Indices (multi-spectral leaf/vegetation segmentation)
This dataset inclue multi-spectral acquisition of vegetation for the conception of new DeepIndices.
1 paper · 1 benchmark
EGO-CH-Gaze (Learning to Detect Attended Objects in Cultural Sites with Gaze Signals and Weak Object Supervision)
To study the problem of weakly supervised attended object detection in cultural sites, we collected and labeled a dataset of egocentric images acquired from subjects visiting a cultural site.
1 paper · 0 benchmarks
EviLOG (Evidential Lidar Occupancy Grid Mapping)
The dataset contains synthetic training, validation and test data for occupancy grid mapping from lidar point clouds.
1 paper · 0 benchmarks
FastZIP Data (FastZIP Dataset and Code)
Structure of code/data folders and how to use them fastzip-code Contains codebase to generate results in fastzip-results folder Individual notebooks contain comments on their functionality/how to use them FastZIP-Resample.ipynb (optional)…
1 paper · 0 benchmarks
The National Health and Nutrition Examination Survey (NHANES) provides data on the health and environmental exposure of the non-institutionalized US population.
1 paper · 0 benchmarks
INSANE Cross-Domain UAV Data Set (Cross-Domain UAV Data Sets with Increased Number of Sensors for developing Advanced and Novel Estimators)
This data set contains over 600GB of multimodal data from a Mars analog mission, including accurate 6DoF outdoor ground truth, indoor-outdoor transitions with continuous cross-domain ground truth, and indoor data with Optitrack…
1 paper · 0 benchmarks
Interactive Gibson is a comprehensive benchmark for training and evaluating Interactive Navigation: robot navigation strategies where physical interaction with objects is allowed and even encouraged to accomplish a task.
1 paper · 0 benchmarks
JAMBO (A Multi-Annotator Image Dataset for Benthic Habitat Classification)
The JAMBO dataset contains 3290 underwater images of the seabed captured by an ROV in temperate waters in the Jammer Bay area off the North West coast of Jutland, Denmark.
1 paper · 0 benchmarks
MAX-60K (Masked Autoencoder for X-ray Fluorescence 60K Dataset)
The dataset for masked autoencoder for X-ray fluorescence (XRF) is a following development after the dataset (Chao et al., 2022).
1 paper · 0 benchmarks
MODIS AOD (imputed) (Pre-processed MODIS AOD and ERA5 data (2003-2022) for North Africa)
Structured atmospheric data for AI/ML Long-term, pre-processed, atmospheric datasets for use in Machine Learning/AI based forecasting.
1 paper · 0 benchmarks
Optimization of pedestrian evacuation in different environments
1 paper · 0 benchmarks
An evaluation dataset for planning with LLM agents
1 paper · 0 benchmarks
PushWorld is an environment with simplistic physics that requires manipulation planning with both movable obstacles and tools.
1 paper · 0 benchmarks
The RoseBlooming dataset is a stage-specific flower dataset for detection.
1 paper · 0 benchmarks
SICKLE (Satellite Imagery for Cropping annotated with Keyparameter LabEls)
The availability of well-curated datasets has driven the success of Machine Learning (ML) models.
1 paper · 1 benchmark
SPAVE-28G (Signal Propagation Analyses in V2X Ecosystems (S.P.A.V.E) at 28 GHz on the NSF POWDER testbed)
Click to add a brief description of the dataset (Markdown and LaTeX enabled).
1 paper · 0 benchmarks
The scene derives from photo-realistic HM3D datasets.
1 paper · 0 benchmarks
The scene derives from photo-realistic MP3D datasets.
1 paper · 0 benchmarks
Sonicverse is a multisensory simulation platform with integrated audio-visual simulation for training household agents that can both see and hear.
1 paper · 0 benchmarks
TERRA-REF (TERRA-REF, An open reference data set from high resolution genomics, phenomics, and imaging sensors)
The ARPA-E funded TERRA-REF project is generating open-access reference datasets for the study of plant sensing, genomics, and phenomics.
1 paper · 0 benchmarks
UIUC Scooping Dataset (Granular Materials Manipulation Dataset with Scooping/Digging/Excavation Action)
Overview: This dataset encompasses a compilation of 6,700 executed scoops (excavations), mapped across a vast spectrum of materials, terrain topography, and compositions.
1 paper · 0 benchmarks
WiFiCam dataset for through-wall imaging based on WiFi channel state information.
1 paper · 0 benchmarks
gComm is a step towards developing a robust platform to foster research in grounded language acquisition in a more challenging and realistic setting.
1 paper · 0 benchmarks
lilGym is a benchmark for language-conditioned reinforcement learning in visual environment based on 2,661 highly-compositional human-written natural language statements grounded in an interactive visual environment.
1 paper · 0 benchmarks
pursuitMW (Multi-agent pursuit in matrix world)
Multi-agent pursuit in matrix world (pursuitMW) is a partially observable Markov game (POMG) between a swarm of pursuers and a swarm of evaders.
1 paper · 0 benchmarks
BEHAVIOR is a benchmark with the 100 household activities that represent a new challenge for embodied AI solutions.
0 papers · 0 benchmarks
The Fields2Benhmark dataset is a collection of 350 agricultural fields in vector format manually selected to test agricultural coverage path planning algorithms.
0 papers · 0 benchmarks
KITTI-360-SR (KITTI-360 modification for Scene Recognition task)
Scene Recognition is a problem, where a set of visible objects must be correctly associated with objects marked on a semantic map - this problem is also sometimes called a Data Association.
0 papers · 0 benchmarks

Paper counts and descriptions are the archive's, frozen 2025-07-28; no citation counts, no stars, no trending. Sorting by "most cited" or "newest" was a live-site feature the archive does not carry.