Methods › Reinforcement Learning › Replay Memory › Experience Replay › Papers where code ran, page 1
Experience Replay
Papers archive 2025-07-28
archive papers tagged: 865 · with a code link: 317 · where Syntology ran a sample: 94 (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (94 of 865 tagged: 86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument)
Syntology We ran code from the paper's repository; we did not isolate this method inside it.
Page 1 of 1: papers 1 to 94 of the 94 tagged papers where Syntology ran at least one harvested sample (86 with a run with no instrument failure, 8 where every run was a failure of Syntology's instrument), newest first by the archive's date (ties by arXiv id). This is a filter on Syntology's record ordered by date only, not a ranking; a run is not a correctness claim. A paper missing from this list is not a recorded non-run: it may have no arXiv id, no harvested code, or only samples that have not run yet.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Gradual Transition from Bellman Optimality Operator to Bellman Operator in Online Reinforcement Learning 6 Jun 2025 · 1 repository · arXiv:2506.05968Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
Trajectory Balance with Asynchrony: Decoupling Exploration and Learning for Fast, Scalable LLM Post-Training 24 Mar 2025 · 1 repository · arXiv:2503.18929Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples)
-
Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning 29 Jan 2025 · 1 repository · arXiv:2501.17827Syntology official (archive's flag): 6 ran · 6 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 2 honoured, 0 violated, 1 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 11 harvested samples)
-
Reinforcement Learning Gradients as Vitamin for Online Finetuning Decision Transformers 31 Oct 2024 · 1 repository · arXiv:2410.24108Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
NetworkGym: Reinforcement Learning Environments for Multi-Access Traffic Management in Network Simulation 30 Oct 2024 · 1 repository · arXiv:2411.04138Syntology official (archive's flag): 6 ran · 15 ran (of which 2 constructed an object rather than computing a result; 12 with no instrument failure: 4 honoured, 0 violated, 8 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 18 harvested samples) · 1 pointer-only (licence)
-
Forgetting, Ignorance or Myopia: Revisiting Key Challenges in Online Continual Learning 28 Sep 2024 · 1 repository · arXiv:2409.19245Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Saturn: Sample-efficient Generative Molecular Design using Memory Manipulation 27 May 2024 · 2 repositories · arXiv:2405.17066Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 1 pointer-only (licence)
-
Continual Learning with Weight Interpolation 5 Apr 2024 · 1 repository · arXiv:2404.04002Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error 7 Mar 2024 · 1 repository · arXiv:2403.04746Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Analyzing and Reducing Catastrophic Forgetting in Parameter Efficient Tuning 29 Feb 2024 · 1 repository · arXiv:2402.18865Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Reconciling Spatial and Temporal Abstractions for Goal Representation 18 Jan 2024 · 1 repository · arXiv:2401.09870Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 11 harvested samples) · 11 pointer-only (licence)
-
Distributional Soft Actor-Critic with Three Refinements 9 Oct 2023 · 2 repositories · arXiv:2310.05858Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Cleanba: A Reproducible and Efficient Distributed Reinforcement Learning Platform 29 Sep 2023 · 1 repository · arXiv:2310.00036Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 8 harvested samples) · 8 pointer-only (licence)
-
Decoupled Prioritized Resampling for Offline RL 8 Jun 2023 · 2 repositories · arXiv:2306.05412Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
For SALE: State-Action Representation Learning for Deep Reinforcement Learning 4 Jun 2023 · 2 repositories · arXiv:2306.02451Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Regularizing Second-Order Influences for Continual Learning 20 Apr 2023 · 1 repository · arXiv:2304.10177Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Synthetic Experience Replay 12 Mar 2023 · 1 repository · arXiv:2303.06614Syntology official (archive's flag): 1 ran · 4 ran (of which 1 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 2 pointer-only (licence)
-
Automatic Noise Filtering with Dynamic Sparse Training in Deep Reinforcement Learning 13 Feb 2023 · 1 repository · arXiv:2302.06548Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Extreme Q-Learning: MaxEnt RL without Entropy 5 Jan 2023 · 4 repositories · arXiv:2301.02328Syntology 8 ran (of which 5 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 3 where Syntology's instrument failed) · 5 unverified (of 13 harvested samples) · 7 pointer-only (licence)
-
The Benefits of Model-Based Generalization in Reinforcement Learning 4 Nov 2022 · 1 repository · arXiv:2211.02222Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Leveraging Demonstrations with Latent Space Priors 26 Oct 2022 · 1 repository · arXiv:2210.14685Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Consistency is the key to further mitigating catastrophic forgetting in continual learning 11 Jul 2022 · 1 repository · arXiv:2207.04998Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Memory Population in Continual Learning via Outlier Elimination 4 Jul 2022 · 1 repository · arXiv:2207.01145Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
USHER: Unbiased Sampling for Hindsight Experience Replay 3 Jul 2022 · 1 repository · arXiv:2207.01115Syntology 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
EnvPool: A Highly Parallel Reinforcement Learning Environment Execution Engine 21 Jun 2022 · 3 repositories · arXiv:2206.10558Syntology official (archive's flag): 3 ran · 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
SYNERgy between SYNaptic consolidation and Experience Replay for general continual learning 8 Jun 2022 · 1 repository · arXiv:2206.04016Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 1 pointer-only (licence)
-
Task-Agnostic Continual Reinforcement Learning: Gaining Insights and Overcoming Challenges 28 May 2022 · 2 repositories · arXiv:2205.14495Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples) · 2 pointer-only (licence)
-
Memory-efficient Reinforcement Learning with Value-based Knowledge Consolidation 22 May 2022 · 1 repository · arXiv:2205.10868Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Dexterous Robotic Manipulation using Deep Reinforcement Learning and Knowledge Transfer for Complex Sparse Reward-based Tasks 19 May 2022 · 1 repository · arXiv:2205.09683Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Continual Sequence Generation with Adaptive Compositional Modules 20 Mar 2022 · 2 repositories · arXiv:2203.10652Syntology official (archive's flag): 14 ran · 14 ran (of which 0 constructed an object rather than computing a result; 14 with no instrument failure: 0 honoured, 0 violated, 14 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 15 harvested samples) · 1 pointer-only (licence)
-
Imitation Learning by State-Only Distribution Matching 9 Feb 2022 · 1 repository · arXiv:2202.04332Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Learning Fast, Learning Slow: A General Continual Learning Method based on Complementary Learning System 29 Jan 2022 · 1 repository · arXiv:2201.12604Syntology official (archive's flag): 4 ran · 4 ran (of which 2 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Adaptively Calibrated Critic Estimates for Deep Reinforcement Learning 24 Nov 2021 · 1 repository · arXiv:2111.12673Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Generalized Decision Transformer for Offline Hindsight Information Matching 19 Nov 2021 · 1 repository · arXiv:2111.10364Syntology official: no sample here; runs from other or unrecorded repositories · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Hit and Lead Discovery with Explorative RL and Fragment-based Molecule Generation 4 Oct 2021 · 0 repositories · arXiv:2110.01219Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Large Batch Experience Replay 4 Oct 2021 · 2 repositories · arXiv:2110.01528Syntology official (archive's flag): 5 ran · 7 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 11 harvested samples)
-
Solving the Real Robot Challenge using Deep Reinforcement Learning 30 Sep 2021 · 2 repositories · arXiv:2109.15233Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples)
-
Trust Region Policy Optimisation in Multi-Agent Reinforcement Learning 23 Sep 2021 · 11 repositories · arXiv:2109.11251Syntology community repositories only · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Safe Deep Reinforcement Learning for Multi-Agent Systems with Continuous Action Spaces 9 Aug 2021 · 1 repository · arXiv:2108.03952Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Adversarial Intrinsic Motivation for Reinforcement Learning 27 May 2021 · 1 repository · arXiv:2105.13345Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 14 harvested samples)
-
New Insights on Reducing Abrupt Representation Change in Online Continual Learning 11 Apr 2021 · 2 repositories · arXiv:2104.05025Syntology official (archive's flag): 3 ran · 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Memory-based Deep Reinforcement Learning for POMDPs 24 Feb 2021 · 1 repository · arXiv:2102.12344Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Adaptive Rational Activations to Boost Deep Reinforcement Learning 18 Feb 2021 · 4 repositories · arXiv:2102.09407Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 3 pointer-only (licence)
-
Tonic: A Deep Reinforcement Learning Library for Fast Prototyping and Benchmarking 15 Nov 2020 · 1 repository · arXiv:2011.07537Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
Rethinking Experience Replay: a Bag of Tricks for Continual Learning 12 Oct 2020 · 3 repositories · arXiv:2010.05595Syntology official (archive's flag): 2 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples)
-
Sample-Efficient Automated Deep Reinforcement Learning 3 Sep 2020 · 1 repository · arXiv:2009.01555Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Revisiting Fundamentals of Experience Replay 13 Jul 2020 · 2 repositories · arXiv:2007.06700Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Self-Supervised Policy Adaptation during Deployment 8 Jul 2020 · 2 repositories · arXiv:2007.04309Syntology 16 ran (of which 9 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 6 where Syntology's instrument failed) · 5 unverified (of 21 harvested samples) · 21 pointer-only (licence)
-
Counterfactual Data Augmentation using Locally Factored Dynamics 6 Jul 2020 · 1 repository · arXiv:2007.02863Syntology 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Experience Replay with Likelihood-free Importance Weights 23 Jun 2020 · 1 repository · arXiv:2006.13169Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Smooth Exploration for Robotic Reinforcement Learning 12 May 2020 · 4 repositories · arXiv:2005.05719Syntology community repositories only · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Reinforcement Learning with Augmented Data 30 Apr 2020 · 2 repositories · arXiv:2004.14990Syntology 21 ran (of which 8 constructed an object rather than computing a result; 9 with no instrument failure: 1 honoured, 0 violated, 8 with no contract checked; 12 where Syntology's instrument failed) · 3 unverified (of 24 harvested samples) · 21 pointer-only (licence)
-
Evolutionary Population Curriculum for Scaling Multi-Agent Reinforcement Learning 23 Mar 2020 · 1 repository · arXiv:2003.10423Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Adversarial Continual Learning 21 Mar 2020 · 1 repository · arXiv:2003.09553Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 1 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 5 pointer-only (licence)
-
FACMAC: Factored Multi-Agent Centralised Policy Gradients 14 Mar 2020 · 3 repositories · arXiv:2003.06709Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 4 harvested samples)
-
Sample Efficient Reinforcement Learning through Learning from Demonstrations in Minecraft 12 Mar 2020 · 3 repositories · arXiv:2003.06066Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
Online Meta-Critic Learning for Off-Policy Actor-Critic Methods 11 Mar 2020 · 1 repository · arXiv:2003.05334Syntology official (archive's flag): 6 ran · 6 ran (of which 6 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 6 samples that ran constructed an object rather than computing a result (of 6 harvested samples) · 6 pointer-only (licence)
-
Reinforcement co-Learning of Deep and Spiking Neural Networks for Energy-Efficient Mapless Navigation with Neuromorphic Hardware 2 Mar 2020 · 1 repository · arXiv:2003.01157Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning 23 Jan 2020 · 4 repositories · arXiv:2001.08726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
SLM Lab: A Comprehensive Benchmark and Modular Software Framework for Reproducible Deep Reinforcement Learning 28 Dec 2019 · 1 repository · arXiv:1912.12482Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Discrete and Continuous Action Representation for Practical RL in Video Games 23 Dec 2019 · 1 repository · arXiv:1912.11077Syntology 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples)
-
TorchBeast: A PyTorch Platform for Distributed RL 8 Oct 2019 · 3 repositories · arXiv:1910.03552Syntology official (archive's flag): 3 ran · 6 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 4 where Syntology's instrument failed) · 3 unverified (of 9 harvested samples) · 3 pointer-only (licence)
-
QuaRL: Quantization for Fast and Environmentally Sustainable Reinforcement Learning 2 Oct 2019 · 1 repository · arXiv:1910.01055Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Online Continual Learning with Maximally Interfered Retrieval 11 Aug 2019 · 1 repository · arXiv:1908.04742Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 13 harvested samples) · 1 pointer-only (licence)
-
Shapley Q-value: A Local Reward Approach to Solve Global Reward Games 11 Jul 2019 · 2 repositories · arXiv:1907.05707Syntology official (archive's flag): 1 ran · 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified; the one sample that ran constructed an object rather than computing a result (of 3 harvested samples) · 3 pointer-only (licence)
-
Proximal Distilled Evolutionary Reinforcement Learning 24 Jun 2019 · 1 repository · arXiv:1906.09807Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Goal-conditioned Imitation Learning 13 Jun 2019 · 1 repository · arXiv:1906.05838Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Boosting Soft Actor-Critic: Emphasizing Recent Experience without Forgetting the Past 10 Jun 2019 · 3 repositories · arXiv:1906.04009Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Improving Exploration in Soft-Actor-Critic with Normalizing Flows Policies 6 Jun 2019 · 1 repository · arXiv:1906.02771Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 3 pointer-only (licence)
-
Episodic Memory in Lifelong Language Learning 3 Jun 2019 · 2 repositories · arXiv:1906.01076Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Collaborative Evolutionary Reinforcement Learning 2 May 2019 · 1 repository · arXiv:1905.00976Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
Deep Reinforcement Learning using Genetic Algorithm for Parameter Optimization 19 Feb 2019 · 2 repositories · arXiv:1905.04100Syntology official (archive's flag): 8 ran · 9 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 13 harvested samples) · 13 pointer-only (licence)
-
Soft Actor-Critic Algorithms and Applications 13 Dec 2018 · 52 repositories · arXiv:1812.05905Syntology official: no sample here; runs from other or unrecorded repositories · 38 ran (of which 0 constructed an object rather than computing a result; 35 with no instrument failure: 4 honoured, 1 violated, 30 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 40 harvested samples) · 5 pointer-only (licence)
-
Off-Policy Deep Reinforcement Learning without Exploration 7 Dec 2018 · 10 repositories · arXiv:1812.02900Syntology community repositories only · 14 ran (of which 12 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 9 pointer-only (licence)
-
Deep Multi-Agent Reinforcement Learning with Relevance Graphs 30 Nov 2018 · 1 repository · arXiv:1811.12557Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples)
-
Learning to Learn without Forgetting by Maximizing Transfer and Minimizing Interference 29 Oct 2018 · 3 repositories · arXiv:1810.11910Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 0 violated, 1 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space 10 Oct 2018 · 5 repositories · arXiv:1810.06394Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Generalizing Across Multi-Objective Reward Functions in Deep Reinforcement Learning 17 Sep 2018 · 1 repository · arXiv:1809.06364Syntology 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Remember and Forget for Experience Replay 16 Jul 2018 · 2 repositories · arXiv:1807.05827Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models 30 May 2018 · 9 repositories · arXiv:1805.12114Syntology community repositories only · 12 ran (of which 5 constructed an object rather than computing a result; 7 with no instrument failure: 1 honoured, 0 violated, 6 with no contract checked; 5 where Syntology's instrument failed) · 7 unverified (of 19 harvested samples) · 17 pointer-only (licence)
-
Distributed Prioritized Experience Replay 2 Mar 2018 · 15 repositories · arXiv:1803.00933Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples) · 3 pointer-only (licence)
-
Addressing Function Approximation Error in Actor-Critic Methods 26 Feb 2018 · 67 repositories · arXiv:1802.09477Syntology community repositories only · 26 ran (of which 0 constructed an object rather than computing a result; 25 with no instrument failure: 1 honoured, 1 violated, 23 with no contract checked; 1 where Syntology's instrument failed) · 10 unverified (of 36 harvested samples) · 21 pointer-only (licence)
-
GEP-PG: Decoupling Exploration and Exploitation in Deep Reinforcement Learning Algorithms 14 Feb 2018 · 1 repository · arXiv:1802.05054Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures 5 Feb 2018 · 24 repositories · arXiv:1802.01561Syntology community repositories only · 16 ran (of which 6 constructed an object rather than computing a result; 10 with no instrument failure: 1 honoured, 1 violated, 8 with no contract checked; 6 where Syntology's instrument failed) · 18 unverified (of 34 harvested samples) · 3 pointer-only (licence)
-
Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor 4 Jan 2018 · 86 repositories · arXiv:1801.01290Syntology community repositories only · 91 ran (of which 61 constructed an object rather than computing a result; 73 with no instrument failure: 3 honoured, 1 violated, 69 with no contract checked; 18 where Syntology's instrument failed) · 57 unverified (of 148 harvested samples) · 66 pointer-only (licence)
-
Overcoming Exploration in Reinforcement Learning with Demonstrations 28 Sep 2017 · 3 repositories · arXiv:1709.10089Syntology 5 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples) · 1 pointer-only (licence)
-
Hindsight Experience Replay 5 Jul 2017 · 28 repositories · arXiv:1707.01495Syntology 16 ran (of which 10 constructed an object rather than computing a result; 11 with no instrument failure: 1 honoured, 0 violated, 10 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 16 harvested samples) · 9 pointer-only (licence)
-
Multi-Agent Actor-Critic for Mixed Cooperative-Competitive Environments 7 Jun 2017 · 86 repositories · arXiv:1706.02275Syntology official: no sample here; runs from other or unrecorded repositories · 75 ran (of which 54 constructed an object rather than computing a result; 68 with no instrument failure: 2 honoured, 0 violated, 66 with no contract checked; 7 where Syntology's instrument failed) · 68 unverified (of 143 harvested samples) · 99 pointer-only (licence)
-
Parameter Space Noise for Exploration 6 Jun 2017 · 10 repositories · arXiv:1706.01905Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Continuous Deep Q-Learning with Model-based Acceleration 2 Mar 2016 · 8 repositories · arXiv:1603.00748Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Prioritized Experience Replay 18 Nov 2015 · 77 repositories · arXiv:1511.05952Syntology 80 ran (of which 62 constructed an object rather than computing a result; 72 with no instrument failure: 4 honoured, 0 violated, 68 with no contract checked; 8 where Syntology's instrument failed) · 31 unverified (of 111 harvested samples) · 43 pointer-only (licence)
-
Deep Reinforcement Learning with Double Q-learning 22 Sep 2015 · 97 repositories · arXiv:1509.06461Syntology 56 ran (of which 38 constructed an object rather than computing a result; 55 with no instrument failure: 0 honoured, 0 violated, 55 with no contract checked; 1 where Syntology's instrument failed) · 50 unverified (of 106 harvested samples) · 57 pointer-only (licence)
-
Continuous control with deep reinforcement learning 9 Sep 2015 · 161 repositories · arXiv:1509.02971Syntology 163 ran (of which 126 constructed an object rather than computing a result; 152 with no instrument failure: 3 honoured, 0 violated, 149 with no contract checked; 11 where Syntology's instrument failed) · 143 unverified (of 306 harvested samples) · 163 pointer-only (licence)
-
Playing Atari with Deep Reinforcement Learning 19 Dec 2013 · 112 repositories · arXiv:1312.5602Syntology 64 ran (of which 24 constructed an object rather than computing a result; 46 with no instrument failure: 5 honoured, 0 violated, 41 with no contract checked; 18 where Syntology's instrument failed) · 53 unverified (of 117 harvested samples) · 56 pointer-only (licence)