{"url":"/dataset/rl-unplugged","name":"RL Unplugged","full_name":"RL Unplugged","description_markdown":"RL Unplugged is suite of benchmarks for offline reinforcement learning. The RL Unplugged is designed around the following considerations: to facilitate ease of use, the datasets are provided with a unified API which makes it easy for the practitioner to work with all data in the suite once a general pipeline has been established. This is a dataset accompanying the paper RL Unplugged: Benchmarks for Offline Reinforcement Learning.\r\n\r\nIn this suite of benchmarks, the authors try to focus on the following problems:\r\n\r\n- High dimensional action spaces, for example the locomotion humanoid domains, there are 56 dimensional actions.\r\n- High dimensional observations.\r\n- Partial observability, observations have egocentric vision.\r\n- Difficulty of exploration, using states of the art algorithms and imitation to generate data for difficult environments.\r\n- Real world challenges.\r\n\r\nSource: [DeepMind](https://github.com/deepmind/deepmind-research/tree/master/rl_unplugged)","description_withheld":null,"homepage":"https://github.com/deepmind/deepmind-research/tree/master/rl_unplugged","introduced_date":"2020-12-01","introduced_date_note":null,"introduced_by":{"paper":"/paper/rl-unplugged-a-collection-of-benchmarks-for","title":"RL Unplugged: A Collection of Benchmarks for Offline Reinforcement Learning","first_author":"Caglar Gulcehre","url":null},"license":{"name":"Apache 2.0","url":"https://www.apache.org/licenses/LICENSE-2.0"},"modalities":[{"name":"Environment","url":"/datasets/modality/environment"}],"tasks":[{"name":"Offline RL","url":"/task/offline-rl","datasets_with_task":"/datasets/task/offline-rl"}],"languages":[],"variants":["RL Unplugged"],"data_loaders":[{"repo":"https://github.com/deepmind/deepmind-research","url":"https://github.com/deepmind/deepmind-research","frameworks":["tf"]}],"num_papers_in_archive":7,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[],"papers_with_a_benchmark_row":[],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":0,"samples_harvested":0,"samples_ran":0,"samples_unverified":0,"pointer_only_for_licence":0,"papers_with_no_sample_that_ran":0,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}