Methods › Reinforcement Learning › Off-Policy TD Control › Q-Learning › Papers, page 16
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,734 · with a code link: 464 · where Syntology ran a sample: 126 (105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (126 of 1,734 tagged: 105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 16 of 18: papers 1,501 to 1,600 of 1,734, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Reinforcement Learning of Markov Decision Processes with Peak Constraints 23 Jan 2019 · 0 repositories · arXiv:1901.07839
-
Understanding Multi-Step Deep Reinforcement Learning: A Systematic Study of the DQN Target 22 Jan 2019 · 1 repository · arXiv:1901.07510
-
A Deep Recurrent Q Network towards Self-adapting Distributed Microservices architecture 13 Jan 2019 · 1 repository · arXiv:1901.04011
-
Deep Reinforcement Learning for Imbalanced Classification 5 Jan 2019 · 3 repositories · arXiv:1901.01379Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
Accelerating Goal-Directed Reinforcement Learning by Model Characterization 4 Jan 2019 · 0 repositories · arXiv:1901.01977
-
A Theoretical Analysis of Deep Q-Learning 1 Jan 2019 · 0 repositories · arXiv:1901.00137
-
Generative Adversarial User Model for Reinforcement Learning Based Recommendation System 27 Dec 2018 · 1 repository · arXiv:1812.10613Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Parallelized Interactive Machine Learning on Autonomous Vehicles 23 Dec 2018 · 0 repositories · arXiv:1812.09724
-
Learning to Navigate the Web 21 Dec 2018 · 0 repositories · arXiv:1812.09195
-
Double Deep Q-Learning for Optimal Execution 17 Dec 2018 · 0 repositories · arXiv:1812.06600
-
Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach 16 Dec 2018 · 2 repositories · arXiv:1812.07394
-
Learning Sharing Behaviors with Arbitrary Numbers of Agents 10 Dec 2018 · 0 repositories · arXiv:1812.04145
-
Off-Policy Deep Reinforcement Learning without Exploration 7 Dec 2018 · 10 repositories · arXiv:1812.02900Syntology community repositories only · 14 ran (of which 12 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 9 pointer-only (licence)
-
Active Deep Q-learning with Demonstration 6 Dec 2018 · 0 repositories · arXiv:1812.02632
-
Bach2Bach: Generating Music Using A Deep Reinforcement Learning Approach 3 Dec 2018 · 0 repositories · arXiv:1812.01060
-
Deep Reinforcement Learning for Intelligent Transportation Systems 3 Dec 2018 · 0 repositories · arXiv:1812.00979
-
Macro action selection with deep reinforcement learning in StarCraft 2 Dec 2018 · 1 repository · arXiv:1812.00336
-
Revisiting the Softmax Bellman Operator: New Benefits and New Perspective 2 Dec 2018 · 2 repositories · arXiv:1812.00456Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Non-delusional Q-learning and value-iteration 1 Dec 2018 · 0 repositories
-
Deep Multi-Agent Reinforcement Learning with Relevance Graphs 30 Nov 2018 · 1 repository · arXiv:1811.12557Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples)
-
Urban Driving with Multi-Objective Deep Reinforcement Learning 21 Nov 2018 · 1 repository · arXiv:1811.08586
-
Reinforcement Learning with A* and a Deep Heuristic 19 Nov 2018 · 2 repositories · arXiv:1811.07745
-
Emergence of Addictive Behaviors in Reinforcement Learning Agents 14 Nov 2018 · 0 repositories · arXiv:1811.05590
-
An initial attempt of combining visual selective attention with deep reinforcement learning 11 Nov 2018 · 0 repositories · arXiv:1811.04407
-
Managing App Install Ad Campaigns in RTB: A Q-Learning Approach 11 Nov 2018 · 0 repositories · arXiv:1811.04475
-
Deep Reinforcement Learning for Green Security Games with Real-Time Information 6 Nov 2018 · 0 repositories · arXiv:1811.02483
-
Reinforcement Learning based Dynamic Model Selection for Short-Term Load Forecasting 5 Nov 2018 · 0 repositories · arXiv:1811.01846
-
Approximate Dynamic Oracle for Dependency Parsing with Reinforcement Learning 1 Nov 2018 · 0 repositories
-
Structure Learning of Deep Neural Networks with Q-Learning 31 Oct 2018 · 0 repositories · arXiv:1810.13155
-
Distributive Dynamic Spectrum Access through Deep Reinforcement Learning: A Reservoir Computing Based Approach 28 Oct 2018 · 0 repositories · arXiv:1810.11758
-
Learning Negotiating Behavior Between Cars in Intersections using Deep Q-Learning 24 Oct 2018 · 0 repositories · arXiv:1810.10469
-
Reconciling λ-Returns with Experience Replay 23 Oct 2018 · 1 repository · arXiv:1810.09967
-
Greedy Actor-Critic: A New Conditional Cross-Entropy Method for Policy Improvement 22 Oct 2018 · 1 repository · arXiv:1810.09103Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
Successor Uncertainties: Exploration and Uncertainty in Temporal Difference Learning 15 Oct 2018 · 2 repositories · arXiv:1810.06530Syntology 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Assessing the Potential of Classical Q-learning in General Game Playing 14 Oct 2018 · 1 repository · arXiv:1810.06078
-
Learning to Sketch with Deep Q Networks and Demonstrated Strokes 14 Oct 2018 · 0 repositories · arXiv:1810.05977
-
Empowerment-driven Exploration using Mutual Information Estimation 11 Oct 2018 · 1 repository · arXiv:1810.05533
-
Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space 10 Oct 2018 · 5 repositories · arXiv:1810.06394Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Deep Quality-Value (DQV) Learning 30 Sep 2018 · 3 repositories · arXiv:1810.00368
-
Generalization and Regularization in DQN 29 Sep 2018 · 1 repository · arXiv:1810.00123
-
Target Transfer Q-Learning and Its Convergence Analysis 21 Sep 2018 · 0 repositories · arXiv:1809.08923
-
Hidden Markov Model Estimation-Based Q-learning for Partially Observable Markov Decision Process 17 Sep 2018 · 0 repositories · arXiv:1809.06401
-
Optimal Matrix Momentum Stochastic Approximation and Applications to Q-learning 17 Sep 2018 · 0 repositories · arXiv:1809.06277
-
Deterministic Implementations for Reproducibility in Deep Reinforcement Learning 15 Sep 2018 · 1 repository · arXiv:1809.05676
-
Sampled Policy Gradient for Learning to Play the Game Agar.io 15 Sep 2018 · 2 repositories · arXiv:1809.05763
-
Towards Better Interpretability in Deep Q-Networks 15 Sep 2018 · 1 repository · arXiv:1809.05630Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Negative Update Intervals in Deep Multi-Agent Reinforcement Learning 13 Sep 2018 · 1 repository · arXiv:1809.05096
-
Coordinated Heterogeneous Distributed Perception based on Latent Space Representation 12 Sep 2018 · 0 repositories · arXiv:1809.04558
-
Learn What Not to Learn: Action Elimination with Deep Reinforcement Learning 6 Sep 2018 · 0 repositories · arXiv:1809.02121
-
Model-Based Regularization for Deep Reinforcement Learning with Transcoder Networks 6 Sep 2018 · 0 repositories · arXiv:1809.01906
-
Directed Exploration in PAC Model-Free Reinforcement Learning 31 Aug 2018 · 0 repositories · arXiv:1808.10552
-
MARL-FWC: Optimal Coordination of Freeway Traffic Control Measures 27 Aug 2018 · 0 repositories · arXiv:1808.09806
-
BlockQNN: Efficient Block-wise Neural Network Architecture Generation 16 Aug 2018 · 2 repositories · arXiv:1808.05584
-
Automatic Derivation Of Formulas Using Reforcement Learning 15 Aug 2018 · 0 repositories · arXiv:1808.04946
-
A Framework for Automated Cellular Network Tuning with Reinforcement Learning 13 Aug 2018 · 2 repositories · arXiv:1808.05140
-
A Reinforcement Learning Approach to Target Tracking in a Camera Network 26 Jul 2018 · 0 repositories · arXiv:1807.10336
-
Accelerated Structure-Aware Reinforcement Learning for Delay-Sensitive Energy Harvesting Wireless Sensors 22 Jul 2018 · 0 repositories · arXiv:1807.08315
-
Discrete linear-complexity reinforcement learning in continuous action spaces for Q-learning algorithms 16 Jul 2018 · 0 repositories · arXiv:1807.06957
-
Is Q-learning Provably Efficient? 10 Jul 2018 · 1 repository · arXiv:1807.03765
-
Video Summarisation by Classification with Deep Reinforcement Learning 9 Jul 2018 · 0 repositories · arXiv:1807.03089
-
Playing against Nature: causal discovery for decision making under uncertainty 3 Jul 2018 · 0 repositories · arXiv:1807.01268
-
Learning to Explore via Meta-Policy Gradient 1 Jul 2018 · 0 repositories
-
Using Reward Machines for High-Level Task Specification and Decomposition in Reinforcement Learning 1 Jul 2018 · 1 repository
-
Many-Goals Reinforcement Learning 22 Jun 2018 · 0 repositories · arXiv:1806.09605
-
Reinforcement Learning using Augmented Neural Networks 20 Jun 2018 · 0 repositories · arXiv:1806.07692
-
Surprising Negative Results for Generative Adversarial Tree Search 15 Jun 2018 · 3 repositories · arXiv:1806.05780
-
Implicit Quantile Networks for Distributional Reinforcement Learning 14 Jun 2018 · 19 repositories · arXiv:1806.06923Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Qualitative Measurements of Policy Discrepancy for Return-Based Deep Q-Network 14 Jun 2018 · 0 repositories · arXiv:1806.06953
-
Automatic formation of the structure of abstract machines in hierarchical reinforcement learning with state clustering 13 Jun 2018 · 0 repositories · arXiv:1806.05292
-
Learning to Search in Long Documents Using Document Structure 9 Jun 2018 · 1 repository · arXiv:1806.03529
-
Fidelity-based Probabilistic Q-learning for Control of Quantum Systems 8 Jun 2018 · 0 repositories · arXiv:1806.03145
-
A Finite Time Analysis of Temporal Difference Learning With Linear Function Approximation 6 Jun 2018 · 0 repositories · arXiv:1806.02450
-
Randomized Value Functions via Multiplicative Normalizing Flows 6 Jun 2018 · 2 repositories · arXiv:1806.02315
-
Hyperparameter Optimization for Tracking With Continuous Deep Q-Learning 1 Jun 2018 · 0 repositories
-
Sample-Efficient Deep Reinforcement Learning via Episodic Backward Update 31 May 2018 · 1 repository · arXiv:1805.12375
-
Depth and nonlinearity induce implicit exploration for RL 29 May 2018 · 0 repositories · arXiv:1805.11711
-
Episodic Memory Deep Q-Networks 19 May 2018 · 0 repositories · arXiv:1805.07603
-
Optimized Computation Offloading Performance in Virtual Edge Computing Systems via Deep Reinforcement Learning 16 May 2018 · 0 repositories · arXiv:1805.06146
-
Advances in Experience Replay 15 May 2018 · 1 repository · arXiv:1805.05536
-
Planning and Learning with Stochastic Action Sets 7 May 2018 · 0 repositories · arXiv:1805.02363
-
A Hybrid Q-Learning Sine-Cosine-based Strategy for Addressing the Combinatorial Test Suite Minimization Problem 27 Apr 2018 · 0 repositories · arXiv:1805.00873
-
Multiagent Soft Q-Learning 25 Apr 2018 · 0 repositories · arXiv:1804.09817
-
Benchmarking projective simulation in navigation problems 23 Apr 2018 · 0 repositories · arXiv:1804.08607
-
Towards Symbolic Reinforcement Learning with Common Sense 23 Apr 2018 · 1 repository · arXiv:1804.08597
-
Reinforced Co-Training 17 Apr 2018 · 0 repositories · arXiv:1804.06035
-
State-Augmentation Transformations for Risk-Sensitive Reinforcement Learning 16 Apr 2018 · 0 repositories · arXiv:1804.05950
-
CytonRL: an Efficient Reinforcement Learning Open-source Toolkit Implemented in C++ 14 Apr 2018 · 1 repository · arXiv:1804.05834
-
MOVI: A Model-Free Approach to Dynamic Fleet Management 13 Apr 2018 · 0 repositories · arXiv:1804.04758
-
Hierarchical Modular Reinforcement Learning Method and Knowledge Acquisition of State-Action Rule for Multi-target Problem 8 Apr 2018 · 0 repositories · arXiv:1804.02698
-
Reinforcement Learning based QoS/QoE-aware Service Function Chaining in Software-Driven 5G Slices 6 Apr 2018 · 0 repositories · arXiv:1804.02099
-
Joint Learning of Interactive Spoken Content Retrieval and Trainable User Simulator 1 Apr 2018 · 0 repositories · arXiv:1804.00318
-
Learning Synergies between Pushing and Grasping with Self-supervised Deep Reinforcement Learning 27 Mar 2018 · 4 repositories · arXiv:1803.09956Syntology official (archive's flag): 3 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 11 harvested samples)
-
Natural Gradient Deep Q-learning 20 Mar 2018 · 0 repositories · arXiv:1803.07482
-
Composable Deep Reinforcement Learning for Robotic Manipulation 19 Mar 2018 · 1 repository · arXiv:1803.06773
-
Learning to Explore with Meta-Policy Gradient 13 Mar 2018 · 0 repositories · arXiv:1803.05044
-
Deep reinforcement learning for time series: playing idealized trading games 11 Mar 2018 · 2 repositories · arXiv:1803.03916
-
Q-CP: Learning Action Values for Cooperative Planning 1 Mar 2018 · 0 repositories · arXiv:1803.00297
-
Variance Reduction Methods for Sublinear Reinforcement Learning 26 Feb 2018 · 0 repositories · arXiv:1802.09184
-
Weighted Double Deep Multiagent Reinforcement Learning in Stochastic Cooperative Environments 23 Feb 2018 · 0 repositories · arXiv:1802.08534
-
A Deep Q-Learning Agent for the L-Game with Variable Batch Training 17 Feb 2018 · 1 repository · arXiv:1802.06225