Methods › Reinforcement Learning › Q-Learning Networks › DQN › Papers, page 5
Deep Q-Network
DQN
Papers archive 2025-07-28
archive papers tagged: 519 · with a code link: 173 · where Syntology ran a sample: 47 (36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (47 of 519 tagged: 36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 5 of 6: papers 401 to 500 of 519, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Benchmarking Batch Deep Reinforcement Learning Algorithms 3 Oct 2019 · 5 repositories · arXiv:1910.01708
-
AI Assisted Annotator using Reinforcement Learning 2 Oct 2019 · 0 repositories · arXiv:1910.02052
-
Deep Reinforcement Learning with Modulated Hebbian plus Q Network Architecture 21 Sep 2019 · 1 repository · arXiv:1909.09902
-
Split Deep Q-Learning for Robust Object Singulation 17 Sep 2019 · 0 repositories · arXiv:1909.08105
-
Reinforcement Learning for Joint Optimization of Multiple Rewards 6 Sep 2019 · 0 repositories · arXiv:1909.02940
-
An Optimistic Perspective on Offline Reinforcement Learning 10 Jul 2019 · 1 repository · arXiv:1907.04543
-
Towards Empathic Deep Q-Learning 26 Jun 2019 · 1 repository · arXiv:1906.10918Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Learning Causal State Representations of Partially Observable Environments 25 Jun 2019 · 0 repositories · arXiv:1906.10437
-
Analysis and Improvement of Adversarial Training in DQN Agents With Adversarially-Guided Exploration (AGE) 3 Jun 2019 · 0 repositories · arXiv:1906.01119
-
RL-Based Method for Benchmarking the Adversarial Resilience and Robustness of Deep Reinforcement Learning Policies 3 Jun 2019 · 0 repositories · arXiv:1906.01110
-
Sequential Triggers for Watermarking of Deep Reinforcement Learning Policies 3 Jun 2019 · 0 repositories · arXiv:1906.01126
-
Learning distant cause and effect using only local and immediate credit assignment 28 May 2019 · 0 repositories · arXiv:1905.11589
-
Prioritized Sequence Experience Replay 25 May 2019 · 0 repositories · arXiv:1905.12726
-
Adaptive Symmetric Reward Noising for Reinforcement Learning 24 May 2019 · 1 repository · arXiv:1905.10144
-
Deep Q-Learning with Q-Matrix Transfer Learning for Novel Fire Evacuation Environment 23 May 2019 · 0 repositories · arXiv:1905.09673
-
Mastering the Game of Sungka from Random Play 17 May 2019 · 1 repository · arXiv:1905.07102
-
Comprehensible Context-driven Text Game Playing 6 May 2019 · 2 repositories · arXiv:1905.02265
-
Beyond Games: Bringing Exploration to Robots in Real-world 1 May 2019 · 0 repositories
-
Inducing Cooperation via Learning to reshape rewards in semi-cooperative multi-agent reinforcement learning 1 May 2019 · 0 repositories
-
Learning agents with prioritization and parameter noise in continuous state and action space 1 May 2019 · 0 repositories
-
Recurrent Experience Replay in Distributed Reinforcement Learning 1 May 2019 · 3 repositories
-
Generative Adversarial Imagination for Sample Efficient Deep Reinforcement Learning 30 Apr 2019 · 0 repositories · arXiv:1904.13255
-
Deep Q Learning Driven CT Pancreas Segmentation with Geometry-Aware U-Net 19 Apr 2019 · 0 repositories · arXiv:1904.09120
-
Personalized Cancer Chemotherapy Schedule: a numerical comparison of performance and robustness in model-based and model-free scheduling methodologies 2 Apr 2019 · 0 repositories · arXiv:1904.01200
-
Lane Change Decision-making through Deep Reinforcement Learning with Rule-based Constraints 30 Mar 2019 · 0 repositories · arXiv:1904.00231
-
DQN with model-based exploration: efficient learning on environments with sparse rewards 22 Mar 2019 · 0 repositories · arXiv:1903.09295
-
Deep Reinforcement Learning with Decorrelation 18 Mar 2019 · 0 repositories · arXiv:1903.07765
-
Reinforcement Learning with Dynamic Boltzmann Softmax Updates 14 Mar 2019 · 1 repository · arXiv:1903.05926
-
Sample-Efficient Model-Free Reinforcement Learning with Off-Policy Critics 11 Mar 2019 · 1 repository · arXiv:1903.04193
-
DeepPool: Distributed Model-free Algorithm for Ride-sharing using Deep Reinforcement Learning 9 Mar 2019 · 0 repositories · arXiv:1903.03882
-
MinAtar: An Atari-Inspired Testbed for Thorough and Reproducible Reinforcement Learning Experiments 7 Mar 2019 · 3 repositories · arXiv:1903.03176Syntology official: harvested, nothing ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Reward Shaping via Meta-Learning 27 Jan 2019 · 0 repositories · arXiv:1901.09330
-
Distillation Strategies for Proximal Policy Optimization 23 Jan 2019 · 0 repositories · arXiv:1901.08128
-
Understanding Multi-Step Deep Reinforcement Learning: A Systematic Study of the DQN Target 22 Jan 2019 · 1 repository · arXiv:1901.07510
-
Deep Reinforcement Learning for Imbalanced Classification 5 Jan 2019 · 3 repositories · arXiv:1901.01379Syntology community repositories only · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 6 harvested samples)
-
A Theoretical Analysis of Deep Q-Learning 1 Jan 2019 · 0 repositories · arXiv:1901.00137
-
Generative Adversarial User Model for Reinforcement Learning Based Recommendation System 27 Dec 2018 · 1 repository · arXiv:1812.10613Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample)
-
Parallelized Interactive Machine Learning on Autonomous Vehicles 23 Dec 2018 · 0 repositories · arXiv:1812.09724
-
Learning to Navigate the Web 21 Dec 2018 · 0 repositories · arXiv:1812.09195
-
Double Deep Q-Learning for Optimal Execution 17 Dec 2018 · 0 repositories · arXiv:1812.06600
-
Decentralized Computation Offloading for Multi-User Mobile Edge Computing: A Deep Reinforcement Learning Approach 16 Dec 2018 · 2 repositories · arXiv:1812.07394
-
Off-Policy Deep Reinforcement Learning without Exploration 7 Dec 2018 · 10 repositories · arXiv:1812.02900Syntology community repositories only · 14 ran (of which 12 constructed an object rather than computing a result; 14 with no instrument failure: 1 honoured, 0 violated, 13 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 14 harvested samples) · 9 pointer-only (licence)
-
Active Deep Q-learning with Demonstration 6 Dec 2018 · 0 repositories · arXiv:1812.02632
-
Bach2Bach: Generating Music Using A Deep Reinforcement Learning Approach 3 Dec 2018 · 0 repositories · arXiv:1812.01060
-
Deep Reinforcement Learning for Intelligent Transportation Systems 3 Dec 2018 · 0 repositories · arXiv:1812.00979
-
Macro action selection with deep reinforcement learning in StarCraft 2 Dec 2018 · 1 repository · arXiv:1812.00336
-
Deep Multi-Agent Reinforcement Learning with Relevance Graphs 30 Nov 2018 · 1 repository · arXiv:1811.12557Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 7 harvested samples)
-
Urban Driving with Multi-Objective Deep Reinforcement Learning 21 Nov 2018 · 1 repository · arXiv:1811.08586
-
An initial attempt of combining visual selective attention with deep reinforcement learning 11 Nov 2018 · 0 repositories · arXiv:1811.04407
-
Reconciling λ-Returns with Experience Replay 23 Oct 2018 · 1 repository · arXiv:1810.09967
-
Successor Uncertainties: Exploration and Uncertainty in Temporal Difference Learning 15 Oct 2018 · 2 repositories · arXiv:1810.06530Syntology 7 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 5 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples) · 7 pointer-only (licence)
-
Empowerment-driven Exploration using Mutual Information Estimation 11 Oct 2018 · 1 repository · arXiv:1810.05533
-
Parametrized Deep Q-Networks Learning: Reinforcement Learning with Discrete-Continuous Hybrid Action Space 10 Oct 2018 · 5 repositories · arXiv:1810.06394Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Generalization and Regularization in DQN 29 Sep 2018 · 1 repository · arXiv:1810.00123
-
Coordinated Heterogeneous Distributed Perception based on Latent Space Representation 12 Sep 2018 · 0 repositories · arXiv:1809.04558
-
Learn What Not to Learn: Action Elimination with Deep Reinforcement Learning 6 Sep 2018 · 0 repositories · arXiv:1809.02121
-
Model-Based Regularization for Deep Reinforcement Learning with Transcoder Networks 6 Sep 2018 · 0 repositories · arXiv:1809.01906
-
Reinforcement Learning using Augmented Neural Networks 20 Jun 2018 · 0 repositories · arXiv:1806.07692
-
Surprising Negative Results for Generative Adversarial Tree Search 15 Jun 2018 · 3 repositories · arXiv:1806.05780
-
Implicit Quantile Networks for Distributional Reinforcement Learning 14 Jun 2018 · 19 repositories · arXiv:1806.06923Syntology 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
Qualitative Measurements of Policy Discrepancy for Return-Based Deep Q-Network 14 Jun 2018 · 0 repositories · arXiv:1806.06953
-
Learning to Search in Long Documents Using Document Structure 9 Jun 2018 · 1 repository · arXiv:1806.03529
-
Randomized Value Functions via Multiplicative Normalizing Flows 6 Jun 2018 · 2 repositories · arXiv:1806.02315
-
Sample-Efficient Deep Reinforcement Learning via Episodic Backward Update 31 May 2018 · 1 repository · arXiv:1805.12375
-
Episodic Memory Deep Q-Networks 19 May 2018 · 0 repositories · arXiv:1805.07603
-
Optimized Computation Offloading Performance in Virtual Edge Computing Systems via Deep Reinforcement Learning 16 May 2018 · 0 repositories · arXiv:1805.06146
-
Advances in Experience Replay 15 May 2018 · 1 repository · arXiv:1805.05536
-
MOVI: A Model-Free Approach to Dynamic Fleet Management 13 Apr 2018 · 0 repositories · arXiv:1804.04758
-
Reinforcement Learning based QoS/QoE-aware Service Function Chaining in Software-Driven 5G Slices 6 Apr 2018 · 0 repositories · arXiv:1804.02099
-
Natural Gradient Deep Q-learning 20 Mar 2018 · 0 repositories · arXiv:1803.07482
-
Weighted Double Deep Multiagent Reinforcement Learning in Stochastic Cooperative Environments 23 Feb 2018 · 0 repositories · arXiv:1802.08534
-
Efficient Exploration through Bayesian Deep Q-Networks 13 Feb 2018 · 1 repository · arXiv:1802.04412
-
Faster Deep Q-learning using Neural Episodic Control 6 Jan 2018 · 0 repositories · arXiv:1801.01968
-
Autonomous Vehicle Fleet Coordination With Deep Reinforcement Learning 1 Jan 2018 · 0 repositories
-
Faster Reinforcement Learning with Expert State Sequences 1 Jan 2018 · 0 repositories
-
PARAMETRIZED DEEP Q-NETWORKS LEARNING: PLAYING ONLINE BATTLE ARENA WITH DISCRETE-CONTINUOUS HYBRID ACTION SPACE 1 Jan 2018 · 1 repository
-
A Deep Policy Inference Q-Network for Multi-Agent Systems 21 Dec 2017 · 0 repositories · arXiv:1712.07893
-
Deep Neuroevolution: Genetic Algorithms Are a Competitive Alternative for Training Deep Neural Networks for Reinforcement Learning 18 Dec 2017 · 12 repositories · arXiv:1712.06567Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 4 pointer-only (licence)
-
Uncertainty Estimates for Efficient Neural Network-based Dialogue Policy Optimisation 30 Nov 2017 · 0 repositories · arXiv:1711.11486
-
A Benchmarking Environment for Reinforcement Learning Based Task Oriented Dialogue Management 29 Nov 2017 · 0 repositories · arXiv:1711.11023
-
Implementing the Deep Q-Network 20 Nov 2017 · 1 repository · arXiv:1711.07478
-
TreeQN and ATreeC: Differentiable Tree-Structured Models for Deep Reinforcement Learning 31 Oct 2017 · 1 repository · arXiv:1710.11417Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 12 harvested samples) · 1 pointer-only (licence)
-
Distributional Reinforcement Learning with Quantile Regression 27 Oct 2017 · 17 repositories · arXiv:1710.10044Syntology 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples) · 2 pointer-only (licence)
-
Rainbow: Combining Improvements in Deep Reinforcement Learning 6 Oct 2017 · 34 repositories · arXiv:1710.02298Syntology 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 1 honoured, 1 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
Deep Reinforcement Learning with Surrogate Agent-Environment Interface 12 Sep 2017 · 0 repositories · arXiv:1709.03942
-
Pre-training Neural Networks with Human Demonstrations for Deep Reinforcement Learning 12 Sep 2017 · 0 repositories · arXiv:1709.04083
-
Formulation of Deep Reinforcement Learning Architecture Toward Autonomous Driving for On-Ramp Merge 7 Sep 2017 · 0 repositories · arXiv:1709.02066
-
LADDER: A Human-Level Bidding Agent for Large-Scale Real-Time Online Auctions 18 Aug 2017 · 0 repositories · arXiv:1708.05565
-
3DCNN-DQN-RNN: A Deep Reinforcement Learning Framework for Semantic Parsing of Large-scale 3D Point Clouds 21 Jul 2017 · 0 repositories · arXiv:1707.06783
-
Noisy Networks for Exploration 30 Jun 2017 · 15 repositories · arXiv:1706.10295Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Parameter Space Noise for Exploration 6 Jun 2017 · 10 repositories · arXiv:1706.01905Syntology 4 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 1 honoured, 1 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
Explaining Transition Systems through Program Induction 23 May 2017 · 0 repositories · arXiv:1705.08320
-
Shallow Updates for Deep Reinforcement Learning 21 May 2017 · 0 repositories · arXiv:1705.07461
-
The Reactor: A fast and sample-efficient Actor-Critic agent for Reinforcement Learning 15 Apr 2017 · 0 repositories · arXiv:1704.04651
-
Deep Q-learning from Demonstrations 12 Apr 2017 · 6 repositories · arXiv:1704.03732
-
Tactics of Adversarial Attack on Deep Reinforcement Learning Agents 8 Mar 2017 · 0 repositories · arXiv:1703.06748
-
Count-Based Exploration with Neural Density Models 3 Mar 2017 · 1 repository · arXiv:1703.01310Syntology 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Sigmoid-Weighted Linear Units for Neural Network Function Approximation in Reinforcement Learning 10 Feb 2017 · 0 repositories · arXiv:1702.03118
-
Autonomous Braking System via Deep Reinforcement Learning 8 Feb 2017 · 2 repositories · arXiv:1702.02302
-
Vulnerability of Deep Reinforcement Learning to Policy Induction Attacks 16 Jan 2017 · 1 repository · arXiv:1701.04143