Methods › Reinforcement Learning › Heuristic Search Algorithms › Monte-Carlo Tree Search › Papers, page 2
Monte-Carlo Tree Search
Papers archive 2025-07-28
archive papers tagged: 166 · with a code link: 62 · where Syntology ran a sample: 12 (10 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (12 of 166 tagged: 10 with a run with no instrument failure, 2 where every run was a failure of Syntology's instrument)
Page 2 of 2: papers 101 to 166 of 166, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Self-play Learning Strategies for Resource Assignment in Open-RAN Networks 3 Mar 2021 · 0 repositories · arXiv:2103.02649
-
Visualizing MuZero Models 25 Feb 2021 · 1 repository · arXiv:2102.12924
-
Combining Off and On-Policy Training in Model-Based Reinforcement Learning 24 Feb 2021 · 0 repositories · arXiv:2102.12194
-
Improving Model-Based Reinforcement Learning with Internal State Representations through Self-Supervision 10 Feb 2021 · 2 repositories · arXiv:2102.05599
-
Deep Learning for General Game Playing with Ludii and Polygames 23 Jan 2021 · 1 repository · arXiv:2101.09562
-
Monte-Carlo Planning and Learning with Language Action Value Estimates 1 Jan 2021 · 0 repositories
-
Playing Nondeterministic Games through Planning with a Learned Model 1 Jan 2021 · 0 repositories
-
Monte-Carlo Graph Search for AlphaZero 20 Dec 2020 · 3 repositories · arXiv:2012.11045
-
Driving-Policy Adaptive Safeguard for Autonomous Vehicles Using Reinforcement Learning 2 Dec 2020 · 0 repositories · arXiv:2012.01010
-
Hierarchical clustering in particle physics through reinforcement learning 16 Nov 2020 · 1 repository · arXiv:2011.08191
-
Critic PI2: Master Continuous Planning via Policy Improvement with Path Integrals and Deep Actor-Critic Reinforcement Learning 13 Nov 2020 · 0 repositories · arXiv:2011.06752
-
On the role of planning in model-based deep reinforcement learning 8 Nov 2020 · 0 repositories · arXiv:2011.04021
-
The Value Equivalence Principle for Model-Based Reinforcement Learning 6 Nov 2020 · 0 repositories · arXiv:2011.03506
-
Interleaving Fast and Slow Decision Making 30 Oct 2020 · 1 repository · arXiv:2010.16244
-
Dream and Search to Control: Latent Space Planning for Continuous Control 19 Oct 2020 · 1 repository · arXiv:2010.09832
-
AlphaZero Based Post-Storm Repair Crew Dispatch for Distribution Grid Restoration 14 Oct 2020 · 0 repositories · arXiv:2010.06764
-
Playing Carcassonne with Monte Carlo Tree Search 27 Sep 2020 · 0 repositories · arXiv:2009.12974
-
Formal Fields: A Framework to Automate Code Generation Across Domains 28 Jul 2020 · 0 repositories · arXiv:2007.14075
-
Monte-Carlo Tree Search as Regularized Policy Optimization 24 Jul 2020 · 3 repositories · arXiv:2007.12509
-
The LoCA Regret: A Consistent Metric to Evaluate Model-Based Behavior in Reinforcement Learning 7 Jul 2020 · 2 repositories · arXiv:2007.03158Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified; the one sample that ran constructed an object rather than computing a result (of 2 harvested samples)
-
Introduction to Behavior Algorithms for Fighting Games 6 Jul 2020 · 0 repositories · arXiv:2007.12586
-
Convex Regularization in Monte-Carlo Tree Search 1 Jul 2020 · 0 repositories · arXiv:2007.00391
-
Practical Massively Parallel Monte-Carlo Tree Search Applied to Molecular Design 18 Jun 2020 · 0 repositories · arXiv:2006.10504
-
Continuous Control for Searching and Planning with a Learned Model 12 Jun 2020 · 0 repositories · arXiv:2006.07430
-
StarCraft II Build Order Optimization using Deep Reinforcement Learning and Monte-Carlo Tree Search 12 Jun 2020 · 0 repositories · arXiv:2006.10525
-
Planning in Markov Decision Processes with Gap-Dependent Sample Complexity 10 Jun 2020 · 0 repositories · arXiv:2006.05879
-
POLY-HOOT: Monte-Carlo Planning in Continuous Space MDPs with Non-Asymptotic Analysis 8 Jun 2020 · 0 repositories · arXiv:2006.04672
-
Manipulating the Distributions of Experience used for Self-Play Learning in Expert Iteration 30 May 2020 · 1 repository · arXiv:2006.00283
-
Unlucky Explorer: A Complete non-Overlapping Map Exploration 28 May 2020 · 0 repositories · arXiv:2005.14156
-
Single-Agent Optimization Through Policy Iteration Using Monte-Carlo Tree Search 22 May 2020 · 0 repositories · arXiv:2005.11335
-
Neural Machine Translation with Monte-Carlo Tree Search 27 Apr 2020 · 1 repository · arXiv:2004.12527
-
Prolog Technology Reinforcement Learning Prover 15 Apr 2020 · 1 repository · arXiv:2004.06997
-
Enhanced Rolling Horizon Evolution Algorithm with Opponent Model Learning: Results for the Fighting Game AI Competition 31 Mar 2020 · 0 repositories · arXiv:2003.13949
-
On Reinforcement Learning for Turn-based Zero-sum Markov Games 25 Feb 2020 · 0 repositories · arXiv:2002.10620
-
Service Selection using Predictive Models and Monte-Carlo Tree Search 12 Feb 2020 · 0 repositories · arXiv:2002.04852
-
Bayesian optimization for backpropagation in Monte-Carlo tree search 25 Jan 2020 · 0 repositories · arXiv:2001.09325
-
Monte-Carlo Tree Search for Policy Optimization 23 Dec 2019 · 0 repositories · arXiv:1912.10648
-
Combining Q-Learning and Search with Amortized Value Estimates 5 Dec 2019 · 0 repositories · arXiv:1912.02807
-
Maximum Entropy Monte-Carlo Planning 1 Dec 2019 · 0 repositories
-
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model 19 Nov 2019 · 18 repositories · arXiv:1911.08265Syntology 43 ran (of which 36 constructed an object rather than computing a result; 43 with no instrument failure: 4 honoured, 0 violated, 39 with no contract checked; 0 where Syntology's instrument failed) · 21 unverified (of 64 harvested samples) · 62 pointer-only (licence)
-
Generalized Mean Estimation in Monte-Carlo Tree Search 1 Nov 2019 · 0 repositories · arXiv:1911.00384
-
Efficient Multivariate Bandit Algorithm with Path Planning 6 Sep 2019 · 0 repositories · arXiv:1909.02705
-
Learning to play the Chess Variant Crazyhouse above World Champion Level with Deep Neural Networks and Human Data 19 Aug 2019 · 3 repositories · arXiv:1908.06660
-
Automated Machine Learning with Monte-Carlo Tree Search 1 Jun 2019 · 2 repositories · arXiv:1906.00170
-
On Value Functions and the Agent-Environment Boundary 30 May 2019 · 0 repositories · arXiv:1905.13341
-
Misleading Authorship Attribution of Source Code using Adversarial Learning 29 May 2019 · 1 repository · arXiv:1905.12386
-
Understanding & Generalizing AlphaGo Zero 1 May 2019 · 0 repositories
-
Monte-Carlo Tree Search for Efficient Visually Guided Rearrangement Planning 23 Apr 2019 · 2 repositories · arXiv:1904.10348Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Structured agents for physical construction 5 Apr 2019 · 0 repositories · arXiv:1904.03177
-
Bayesian Reinforcement Learning in Factored POMDPs 14 Nov 2018 · 0 repositories · arXiv:1811.05612
-
Improving Hearthstone AI by Combining MCTS and Supervised Learning Algorithms 14 Aug 2018 · 0 repositories · arXiv:1808.04794
-
Hierarchical Reinforcement Learning for Zero-shot Generalization with Subtask Dependencies 19 Jul 2018 · 1 repository · arXiv:1807.07665Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Feedback-Based Tree Search for Reinforcement Learning 15 May 2018 · 0 repositories · arXiv:1805.05935
-
Active Reinforcement Learning with Monte-Carlo Tree Search 13 Mar 2018 · 0 repositories · arXiv:1803.04926
-
Q-CP: Learning Action Values for Cooperative Planning 1 Mar 2018 · 0 repositories · arXiv:1803.00297
-
Latent forward model for Real-time Strategy game planning with incomplete information 1 Jan 2018 · 0 repositories
-
Neural Task Graph Execution 1 Jan 2018 · 0 repositories
-
Adaptive Motion Gaming AI for Health Promotion 4 Apr 2017 · 0 repositories · arXiv:1704.00961
-
Generalised Discount Functions applied to a Monte-Carlo AImu Implementation 3 Mar 2017 · 1 repository · arXiv:1703.01358
-
Monte Carlo Planning method estimates planning horizons during interactive social exchange 12 Feb 2015 · 0 repositories · arXiv:1502.03696
-
Move Evaluation in Go Using Deep Convolutional Neural Networks 20 Dec 2014 · 1 repository · arXiv:1412.6564
-
Monte-Carlo Planning: Theoretically Fast Convergence Meets Practical Efficiency 26 Sep 2013 · 0 repositories · arXiv:1309.6828
-
Geiringer Theorems: From Population Genetics to Computational Intelligence, Memory Evolutive Systems and Hebbian Learning 11 May 2013 · 0 repositories · arXiv:1305.2504
-
Monte-Carlo Planning in Large POMDPs 1 Dec 2010 · 0 repositories
-
Feature Selection as a One-Player Game 17 May 2010 · 0 repositories
-
A Monte Carlo AIXI Approximation 4 Sep 2009 · 2 repositories · arXiv:0909.0801