Browse State-of-the-Art › NetHack
NetHack
22 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Mean in-game score over 1000 episodes with random seeds not seen during training. See https://arxiv.org/abs/2006.13760 (Section 2.4 Evaluation Protocol) for details.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
1 subtask in the archive's task tree.
Most implemented papers archive 2025-07-28
22 shown of 22 papers with code (28 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
24 Jun 2020 3 repositories listed Syntology ran 1 of 11 samples · 10 unverified · 3 pointer-only (licence)Here, we present the NetHack Learning Environment (NLE), a scalable, procedurally generated, stochastic, rich, and challenging environment for RL research based on the popular single-player terminal-based roguelike…
-
23 Jul 2022 2 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedIn this paper, we investigate how skills can be incorporated into the training of reinforcement learning (RL) agents in complex environments with large state-action spaces and sparse rewards.
-
19 Oct 2021 2 repositories listedIn this work, we present CORA, a platform for Continual Reinforcement Learning Agents that provides benchmarks, baselines, and metrics in a single code package.
-
15 Dec 2020 2 repositories listedIn this paper, we analyze the pros and cons of each method and propose the regulated difference of inverse visitation counts as a simple but effective criterion for IR.
-
18 Nov 2024 1 repository listedSyllabus provides a universal API for curriculum learning algorithms, implementations of popular curriculum learning methods, and infrastructure for easily integrating them with distributed training code written in…
-
30 Oct 2024 1 repository listed Syntology ran 0 of 12 samples · 12 unverified · 12 pointer-only (licence)Our approach annotates the agent's collected experience via an asynchronous LLM server, which is then distilled into an intrinsic reward model.
-
11 Jun 2024 1 repository listedYou have an environment, a model, and a reinforcement learning library that are designed to work together but don't.
-
1 Mar 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedIn contrast, agents tested in dynamic robot environments face limitations due to simplistic environments with only a few objects and interactions.
-
26 Feb 2024 1 repository listed Syntology ran 2 of 16 samples · 14 unverifiedEither they are too slow for meaningful research to be performed without enormous computational resources, like Crafter, NetHack and Minecraft, or they are not complex enough to pose a significant challenge, like…
-
5 Feb 2024 1 repository listed Syntology ran 2 of 3 samples · 1 unverifiedWe evaluate our method in the classic videogame NetHack and the text environment ScienceWorld to demonstrate SSO's ability to optimize a set of skills and perform in-context policy improvement.
-
5 Feb 2024 1 repository listedFine-tuning is a widespread technique that allows practitioners to transfer pre-trained capabilities, as recently showcased by the successful applications of foundation models.
-
12 Dec 2023 1 repository listed Syntology ran 4 of 7 samples · 3 unverifiedOn NetHack, an unsolved video game that requires long-horizon reasoning for decision-making, LMs tuned with diff history match state-of-the-art performance for neural agents while needing 1800x fewer training examples…
-
29 Sep 2023 1 repository listed Syntology ran 0 of 5 samples · 5 unverified · 5 pointer-only (licence)Exploring rich environments and evaluating one's actions without prior knowledge is immensely challenging.
-
18 Jul 2023 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Inspired by recent work in Natural Language Processing (NLP) where "scaling up" has resulted in increasingly more capable LLMs, we investigate whether carefully scaling up model and data size can bring similar…
-
17 Jul 2023 1 repository listedIn the last few decades we have witnessed a significant development in Artificial Intelligence (AI) thanks to the availability of a variety of testbeds, mostly based on simulated environments and video games.
-
14 Jun 2023 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)NetHack is known as the frontier of reinforcement learning research where learning-based methods still need to catch up to rule-based solutions.
-
1 Nov 2022 1 repository listed Syntology ran 1 of 3 samples · 2 unverified · 3 pointer-only (licence)Recent breakthroughs in the development of agents to solve challenging sequential decision making problems such as Go, StarCraft, or DOTA, have relied on both simulated environments and large-scale datasets.
-
30 Sep 2022 1 repository listed Syntology ran 2 of 8 samples · 6 unverified · 4 pointer-only (licence)Recent work has shown that augmenting environments with language descriptions improves policy learning.
-
22 Mar 2022 1 repository listed Syntology ran 0 of 2 samples · 2 unverifiedIn this report, we summarize the takeaways from the first NeurIPS 2021 NetHack Challenge.
-
1 Dec 2021 1 repository listedWe analyze NovelD thoroughly in MiniGrid and found that empirically it helps the agent explore the environment more uniformly with a focus on exploring beyond the boundary.
-
20 Oct 2021 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 2 pointer-only (licence)We hope SILG enables the community to quickly identify new methodologies for language grounding that generalize to a diverse set of environments and their associated challenges.
-
27 Sep 2021 1 repository listedBy leveraging the full set of entities and environment dynamics from NetHack, one of the richest grid-based video games, MiniHack allows designing custom RL testbeds that are fast and convenient to use.
Syntology lines on 14 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections