Methods › Reinforcement Learning › Q-Learning Networks › DQN › Papers, page 4
Deep Q-Network
DQN
Papers archive 2025-07-28
archive papers tagged: 519 · with a code link: 173 · where Syntology ran a sample: 47 (36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (47 of 519 tagged: 36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 4 of 6: papers 301 to 400 of 519, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Deep Reinforcement Agent for Scheduling in HPC 11 Feb 2021 · 1 repository · arXiv:2102.06243
-
A review of motion planning algorithms for intelligent robotics 4 Feb 2021 · 0 repositories · arXiv:2102.02376
-
Benchmarking Perturbation-based Saliency Maps for Explaining Atari Agents 18 Jan 2021 · 1 repository · arXiv:2101.07312
-
Action Priors for Large Action Spaces in Robotics 11 Jan 2021 · 1 repository · arXiv:2101.04178
-
Evolving Reinforcement Learning Algorithms 8 Jan 2021 · 5 repositories · arXiv:2101.03958
-
Reinforced Imitative Graph Representation Learning for Mobile User Profiling: An Adversarial Training Perspective 7 Jan 2021 · 0 repositories · arXiv:2101.02634
-
Reinforcement Learning with Latent Flow 6 Jan 2021 · 2 repositories · arXiv:2101.01857Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
A novel policy for pre-trained Deep Reinforcement Learning for Speech Emotion Recognition 4 Jan 2021 · 1 repository · arXiv:2101.00738
-
Learning to Search for Fast Maximum Common Subgraph Detection 1 Jan 2021 · 0 repositories
-
PAC-Bayesian Randomized Value Function with Informative Prior 1 Jan 2021 · 0 repositories
-
Preventing Value Function Collapse in Ensemble Q-Learning by Maximizing Representation Diversity 1 Jan 2021 · 0 repositories
-
Weighted Bellman Backups for Improved Signal-to-Noise in Q-Updates 1 Jan 2021 · 0 repositories
-
Disentangled Planning and Control in Vision Based Robotics via Reward Machines 28 Dec 2020 · 0 repositories · arXiv:2012.14464
-
A State Representation Dueling Network for Deep Reinforcement Learning 24 Dec 2020 · 0 repositories
-
Deploying Reinforcement Learning in Water Transport 14 Dec 2020 · 0 repositories
-
Learn to Play Tetris with Deep Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Mobile Robots Autonomous Exploration with Reinforcement Learning 14 Dec 2020 · 0 repositories
-
Policy Gradient RL Algorithms as Directed Acyclic Graphs 14 Dec 2020 · 1 repository · arXiv:2012.07763
-
Self-correcting Q-Learning 2 Dec 2020 · 0 repositories · arXiv:2012.01100
-
A new convergent variant of Q-learning with linear function approximation 1 Dec 2020 · 0 repositories
-
Predictive PER: Balancing Priority and Diversity towards Stable Deep Reinforcement Learning 26 Nov 2020 · 0 repositories · arXiv:2011.13093
-
Revisiting Rainbow: Promoting more Insightful and Inclusive Deep Reinforcement Learning Research 20 Nov 2020 · 2 repositories · arXiv:2011.14826
-
Leveraging the Variance of Return Sequences for Exploration Policy 17 Nov 2020 · 0 repositories · arXiv:2011.08649
-
Optimizing Large-Scale Fleet Management on a Road Network using Multi-Agent Deep Reinforcement Learning with Graph Neural Network 12 Nov 2020 · 1 repository · arXiv:2011.06175
-
Hamilton-Jacobi Deep Q-Learning for Deterministic Continuous-Time Systems with Lipschitz Continuous Controls 27 Oct 2020 · 1 repository · arXiv:2010.14087
-
Adversarial Attacks on Deep Algorithmic Trading Policies 22 Oct 2020 · 0 repositories · arXiv:2010.11388
-
A Weighted Heterogeneous Graph Based Dialogue System 21 Oct 2020 · 0 repositories · arXiv:2010.10699
-
Chance-Constrained Control with Lexicographic Deep Reinforcement Learning 19 Oct 2020 · 0 repositories · arXiv:2010.09468
-
Connections between Relational Event Model and Inverse Reinforcement Learning for Characterizing Group Interaction Sequences 19 Oct 2020 · 1 repository · arXiv:2010.09810
-
Multi-Agent Reinforcement Learning in NOMA-aided UAV Networks for Cellular Offloading 18 Oct 2020 · 0 repositories · arXiv:2010.09094
-
NOMA in UAV-aided cellular offloading: A machine learning approach 18 Oct 2020 · 0 repositories · arXiv:2011.14776
-
Machine Learning Empowered Trajectory and Passive Beamforming Design in UAV-RIS Wireless Networks 6 Oct 2020 · 0 repositories · arXiv:2010.02749
-
Bayesian Meta-reinforcement Learning for Traffic Signal Control 1 Oct 2020 · 0 repositories · arXiv:2010.00163
-
Strategy and Benchmark for Converting Deep Q-Networks to Event-Driven Spiking Neural Networks 30 Sep 2020 · 0 repositories · arXiv:2009.14456
-
Lineage Evolution Reinforcement Learning 26 Sep 2020 · 0 repositories · arXiv:2010.14616
-
A New Approach for Tactical Decision Making in Lane Changing: Sample Efficient Deep Q Learning with a Safety Feedback Reward 24 Sep 2020 · 0 repositories · arXiv:2009.11905
-
Tactical Decision Making for Emergency Vehicles Based on A Combinational Learning Method 9 Sep 2020 · 0 repositories · arXiv:2009.04203
-
DRLE: Decentralized Reinforcement Learning at the Edge for Traffic Light Control in the IoV 3 Sep 2020 · 1 repository · arXiv:2009.01502
-
An adaptive synchronization approach for weights of deep reinforcement learning 16 Aug 2020 · 0 repositories · arXiv:2008.06973
-
Chrome Dino Run using Reinforcement Learning 15 Aug 2020 · 0 repositories · arXiv:2008.06799
-
Reinforcement Learning with Quantum Variational Circuits 15 Aug 2020 · 2 repositories · arXiv:2008.07524
-
Generalized Radio Environment Monitoring for Next Generation Wireless Networks 14 Aug 2020 · 0 repositories · arXiv:2008.06203
-
Convex Q-Learning, Part 1: Deterministic Optimal Control 8 Aug 2020 · 0 repositories · arXiv:2008.03559
-
Deep Q-Network Based Multi-agent Reinforcement Learning with Binary Action Agents 6 Aug 2020 · 0 repositories · arXiv:2008.04109
-
Deep Reinforcement Learning for Dynamic Spectrum Sensing and Aggregation in Multi-Channel Wireless Networks 28 Jul 2020 · 0 repositories · arXiv:2007.13965
-
Mixture of Step Returns in Bootstrapped DQN 16 Jul 2020 · 0 repositories · arXiv:2007.08229
-
Analysis of Q-learning with Adaptation and Momentum Restart for Gradient Descent 15 Jul 2020 · 0 repositories · arXiv:2007.07422
-
Simulating multi-exit evacuation using deep reinforcement learning 11 Jul 2020 · 0 repositories · arXiv:2007.05783
-
SUNRISE: A Simple Unified Framework for Ensemble Learning in Deep Reinforcement Learning 9 Jul 2020 · 1 repository · arXiv:2007.04938
-
Auto-MAP: A DQN Framework for Exploring Distributed Execution Plans for DNN Workloads 8 Jul 2020 · 0 repositories · arXiv:2007.04069
-
Cognitive Radio Network Throughput Maximization with Deep Reinforcement Learning 7 Jul 2020 · 0 repositories · arXiv:2007.03165
-
Decentralized Deep Reinforcement Learning for Network Level Traffic Signal Control 2 Jul 2020 · 0 repositories · arXiv:2007.03433
-
Some approaches used to overcome overestimation in Deep Reinforcement Learning algorithms 25 Jun 2020 · 0 repositories · arXiv:2006.14167
-
Preventing Value Function Collapse in Ensemble Q-Learning by Maximizing Representation Diversity 24 Jun 2020 · 0 repositories · arXiv:2006.13823
-
RL Unplugged: A Suite of Benchmarks for Offline Reinforcement Learning 24 Jun 2020 · 2 repositories · arXiv:2006.13888
-
Efficient Ridesharing Dispatch Using Multi-Agent Reinforcement Learning 18 Jun 2020 · 1 repository · arXiv:2006.10897
-
Interaction Networks: Using a Reinforcement Learner to train other Machine Learning algorithms 15 Jun 2020 · 0 repositories · arXiv:2006.08457
-
Balancing a CartPole System with Reinforcement Learning -- A Tutorial 8 Jun 2020 · 0 repositories · arXiv:2006.04938
-
Conservative Q-Learning for Offline Reinforcement Learning 8 Jun 2020 · 18 repositories · arXiv:2006.04779Syntology community repositories only · 31 ran (of which 5 constructed an object rather than computing a result; 28 with no instrument failure: 2 honoured, 0 violated, 26 with no contract checked; 3 where Syntology's instrument failed) · 3 unverified (of 34 harvested samples) · 19 pointer-only (licence)
-
Acme: A Research Framework for Distributed Reinforcement Learning 1 Jun 2020 · 5 repositories · arXiv:2006.00979
-
Learning to Charge RF-Energy Harvesting Devices in WiFi Networks 25 May 2020 · 0 repositories · arXiv:2005.12022
-
Prototypical Q Networks for Automatic Conversational Diagnosis and Few-Shot New Disease Adaption 19 May 2020 · 0 repositories · arXiv:2005.11153
-
Local and Global Explanations of Agent Behavior: Integrating Strategy Summaries with Saliency Maps 18 May 2020 · 1 repository · arXiv:2005.08874Syntology official (archive's flag): 9 ran · 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 6 unverified (of 15 harvested samples)
-
Risk-Aware High-level Decisions for Automated Driving at Occluded Intersections with Reinforcement Learning 9 Apr 2020 · 0 repositories · arXiv:2004.04450
-
An Application of Deep Reinforcement Learning to Algorithmic Trading 7 Apr 2020 · 1 repository · arXiv:2004.06627
-
Uniform State Abstraction For Reinforcement Learning 6 Apr 2020 · 0 repositories · arXiv:2004.02919
-
Importance of using appropriate baselines for evaluation of data-efficiency in deep reinforcement learning for Atari 23 Mar 2020 · 0 repositories · arXiv:2003.10181
-
Deep Constrained Q-learning 20 Mar 2020 · 0 repositories · arXiv:2003.09398
-
Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations 19 Mar 2020 · 4 repositories · arXiv:2003.08938
-
Simultaneous Navigation and Radio Mapping for Cellular-Connected UAV with Deep Reinforcement Learning 17 Mar 2020 · 1 repository · arXiv:2003.07574
-
Application of Deep Q-Network in Portfolio Management 13 Mar 2020 · 0 repositories · arXiv:2003.06365
-
Dynamic Experience Replay 4 Mar 2020 · 0 repositories · arXiv:2003.02372
-
Contention Window Optimization in IEEE 802.11ax Networks with Deep Reinforcement Learning 3 Mar 2020 · 1 repository · arXiv:2003.01492
-
Optimistic Exploration even with a Pessimistic Initialisation 26 Feb 2020 · 1 repository · arXiv:2002.12174Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Disentangling Controllable Object through Video Prediction Improves Visual Reinforcement Learning 21 Feb 2020 · 0 repositories · arXiv:2002.09136
-
Langevin DQN 17 Feb 2020 · 2 repositories · arXiv:2002.07282
-
A Multimodal Dialogue System for Conversational Image Editing 16 Feb 2020 · 0 repositories · arXiv:2002.06484
-
Reinforced active learning for image segmentation 16 Feb 2020 · 1 repository · arXiv:2002.06583
-
Fast Reinforcement Learning for Anti-jamming Communications 13 Feb 2020 · 0 repositories · arXiv:2002.05364
-
Safe Wasserstein Constrained Deep Q-Learning 7 Feb 2020 · 0 repositories · arXiv:2002.03016
-
Deep Radial-Basis Value Functions for Continuous Control 5 Feb 2020 · 0 repositories · arXiv:2002.01883
-
Interpretable End-to-end Urban Autonomous Driving with Latent Deep Reinforcement Learning 23 Jan 2020 · 4 repositories · arXiv:2001.08726Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples)
-
Deep Interactive Reinforcement Learning for Path Following of Autonomous Underwater Vehicle 10 Jan 2020 · 0 repositories · arXiv:2001.03359
-
Deep Randomized Least Squares Value Iteration 1 Jan 2020 · 0 repositories
-
SLM Lab: A Comprehensive Benchmark and Modular Software Framework for Reproducible Deep Reinforcement Learning 28 Dec 2019 · 1 repository · arXiv:1912.12482Syntology official (archive's flag): 10 ran · 10 ran (of which 0 constructed an object rather than computing a result; 10 with no instrument failure: 0 honoured, 0 violated, 10 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 13 harvested samples)
-
Exploiting the potential of deep reinforcement learning for classification tasks in high-dimensional and unstructured data 20 Dec 2019 · 0 repositories · arXiv:1912.09595
-
Soft Q Network 20 Dec 2019 · 0 repositories · arXiv:1912.10891
-
Learning Sparse Representations Incrementally in Deep Reinforcement Learning 9 Dec 2019 · 0 repositories · arXiv:1912.04002
-
Reconciling λ-Returns with Experience Replay 1 Dec 2019 · 1 repository
-
Placement Optimization of Aerial Base Stations with Deep Reinforcement Learning 19 Nov 2019 · 0 repositories · arXiv:1911.08111
-
Minimalistic Attacks: How Little it Takes to Fool a Deep Reinforcement Learning Policy 10 Nov 2019 · 0 repositories · arXiv:1911.03849
-
An End-to-End Deep RL Framework for Task Arrangement in Crowdsourcing Platforms 4 Nov 2019 · 0 repositories · arXiv:1911.01030
-
Task-Oriented Language Grounding for Language Input with Multiple Sub-Goals of Non-Linear Order 27 Oct 2019 · 1 repository · arXiv:1910.12354
-
Momentum in Reinforcement Learning 21 Oct 2019 · 0 repositories · arXiv:1910.09322
-
Resource Allocation in Mobility-Aware Federated Learning Networks: A Deep Reinforcement Learning Approach 21 Oct 2019 · 0 repositories · arXiv:1910.09172
-
Reverse Experience Replay 19 Oct 2019 · 0 repositories · arXiv:1910.08780
-
Learning Visual Affordances with Target-Orientated Deep Q-Network to Grasp Objects by Harnessing Environmental Fixtures 9 Oct 2019 · 0 repositories · arXiv:1910.03781
-
Multi-step Greedy Reinforcement Learning Algorithms 7 Oct 2019 · 0 repositories · arXiv:1910.02919
-
Deep Q-Network for Angry Birds 4 Oct 2019 · 1 repository · arXiv:1910.01806
-
I'm sorry Dave, I'm afraid I can't do that, Deep Q-learning from forbidden action 4 Oct 2019 · 0 repositories · arXiv:1910.02078