Browse State-of-the-Art › Sokoban
Sokoban
32 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 32 papers with code (61 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
27 Mar 2019 5 repositories listedWe compare them on three different game level generation problems: Binary, Zelda, and Sokoban.
-
22 Jul 2024 2 repositories listed Syntology ran 4 of 4 samples · 0 unverified · 1 pointer-only (licence)How a neural network (NN) generalizes to novel situations depends on whether it has learned to select actions heuristically or via a planning process.
-
26 May 2023 2 repositories listedLevin Tree Search (LTS) is a search algorithm that makes use of a policy (a probability distribution over actions) and comes with a theoretical guarantee on the number of expansions before reaching a goal node,…
-
12 Sep 2021 2 repositories listed Syntology ran 2 of 13 samples · 11 unverifiedWe present a method of generating diverse collections of neural cellular automata (NCA) to design video game levels.
-
23 Jun 2019 2 repositories listedThis problem is central to inductive general game playing (IGGP).
-
6 Oct 2018 2 repositories listedBeing able to reach any desired location in the environment can be a valuable asset for an agent.
-
13 Feb 2018 2 repositories listedThey are most typically solved by tree search algorithms that simulate ahead into the future, evaluate future states, and back-up those evaluations to the root of a search tree.
-
11 Jun 2025 1 repository listedEach layer has its own plan representation and value function, increasing search depth.
-
6 Apr 2025 1 repository listedWe introduce a novel hierarchical reinforcement learning (HRL) framework that performs top-down recursive planning via learned subgoals, successfully applied to the complex combinatorial puzzle game Sokoban.
-
5 Dec 2024 1 repository listedSuch vision-in-the-chain reasoning paradigm is more aligned with the needs of multimodal agents, while being rarely evaluated.
-
14 Sep 2024 1 repository listedWhile model-based reinforcement learning methods learn world models that can then be used for planning, such approaches are limited by errors that accumulate when the model is applied across many timesteps as well as…
-
13 Jul 2024 1 repository listedCombining Large Language Models (LLMs) with heuristic search algorithms like A* holds the promise of enhanced LLM reasoning and scalable inference.
-
13 Jun 2024 1 repository listedRecently, procedural content generation has exhibited considerable advancements in the domain of 2D game level generation such as Super Mario Bros.
-
21 Feb 2024 1 repository listed Syntology ran 13 of 14 samples · 1 unverified · 14 pointer-only (licence)We fine tune this model to obtain a Searchformer, a Transformer model that optimally solves previously unseen Sokoban puzzles 93.
-
2 Oct 2023 1 repository listedMany puzzle video games, like Sokoban, involve moving some agent in a maze.
-
27 Jul 2023 1 repository listed Syntology ran 2 of 2 samples · 0 unverifiedThis approach eliminates the need for handcrafted planning algorithms by enabling the agent to learn how to plan autonomously and allows for easy interpretation of the agent's plan with visualization.
-
29 Sep 2022 1 repository listedA level generator is a tool that generates game levels from noise.
-
1 Jun 2022 1 repository listedComplex reasoning problems contain states that vary in the computational cost required to determine a good action plan.
-
3 Feb 2022 1 repository listedDepending upon the smoothness of the action-value function, one approach to overcoming this issue is through online learning, where information is interpolated among similar states; Policy Gradient Search provides a…
-
25 Aug 2021 1 repository listedIn this paper, we implement kSubS using a transformer-based subgoal module coupled with the classical best-first search framework.
-
30 Jun 2021 1 repository listedCurrent domain-independent, classical planners require symbolic models of the problem domain and instance as input, resulting in a knowledge acquisition bottleneck.
-
21 Mar 2021 1 repository listed Syntology ran 3 of 3 samples · 0 unverifiedLevinTS is guided by a policy and provides guarantees on the number of search steps that relate to the quality of the policy, but it does not make use of a heuristic function.
-
9 Mar 2021 1 repository listedIn this paper, we present our findings with using action model learning (AML), in which an action model is learned given data in the form of a play trace, to learn a player model in a domain-agnostic manner.
-
11 Sep 2020 1 repository listedTo encourage progress towards this goal we introduce a set of physically embedded planning problems and make them publicly available.
-
6 Aug 2020 1 repository listedThis paper introduces RL Brush, a level-editing tool for tile-based games designed for mixed-initiative co-creation.
-
4 Jun 2020 1 repository listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Despite significant progress in general AI planning, certain domains remain out of reach of current AI planning systems.
-
17 May 2020 1 repository listedThis paper introduces a new system to design constructive level generators by searching the space of constructive level generators defined by Marahel language.
-
19 Dec 2019 1 repository listedThe former manifests itself through the use of value function, while the latter is powered by a tree search planner.
-
25 Sep 2019 1 repository listedNotably, our method performs well in environments with sparse rewards where standard TD(1) backups fail.
-
1 Sep 2019 1 repository listedThis paper examines learning approaches for forward models based on local cell transition functions.
Syntology lines on 6 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections