Methods › Reinforcement Learning › Q-Learning Networks › DQN › Papers, page 3
Deep Q-Network
DQN
Papers archive 2025-07-28
archive papers tagged: 519 · with a code link: 173 · where Syntology ran a sample: 47 (36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (47 of 519 tagged: 36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 3 of 6: papers 201 to 300 of 519, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Enforcing KL Regularization in General Tsallis Entropy Reinforcement Learning via Advantage Learning 16 May 2022 · 0 repositories · arXiv:2205.07885
-
Efficient Off-Policy Reinforcement Learning via Brain-Inspired Computing 14 May 2022 · 0 repositories · arXiv:2205.06978
-
Characterizing the Action-Generalization Gap in Deep Q-Learning 11 May 2022 · 0 repositories · arXiv:2205.05588
-
EPiDA: An Easy Plug-in Data Augmentation Framework for High Performance Text Classification 24 Apr 2022 · 1 repository · arXiv:2204.11205Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 6 harvested samples) · 2 pointer-only (licence)
-
Graph Neural Network based Agent in Google Research Football 23 Apr 2022 · 0 repositories · arXiv:2204.11142
-
Learning how to Interact with a Complex Interface using Hierarchical Reinforcement Learning 21 Apr 2022 · 0 repositories · arXiv:2204.10374
-
Reinforcement Re-ranking with 2D Grid-based Recommendation Panels 11 Apr 2022 · 0 repositories · arXiv:2204.04954
-
DouZero+: Improving DouDizhu AI by Opponent Modeling and Coach-guided Learning 6 Apr 2022 · 1 repository · arXiv:2204.02558
-
REM: Routing Entropy Minimization for Capsule Networks 4 Apr 2022 · 0 repositories · arXiv:2204.01298
-
Hysteresis-Based RL: Robustifying Reinforcement Learning-based Control Policies via Hybrid Control 1 Apr 2022 · 2 repositories · arXiv:2204.00654
-
Investigating the Properties of Neural Network Representations in Reinforcement Learning 30 Mar 2022 · 0 repositories · arXiv:2203.15955
-
MERLIN -- Malware Evasion with Reinforcement LearnINg 24 Mar 2022 · 0 repositories · arXiv:2203.12980
-
Does DQN really learn? Exploring adversarial training schemes in Pong 20 Mar 2022 · 0 repositories · arXiv:2203.10614
-
Random Ensemble Reinforcement Learning for Traffic Signal Control 10 Mar 2022 · 0 repositories · arXiv:2203.05961
-
Deep Reinforcement Learning based Model-free On-line Dynamic Multi-Microgrid Formation to Enhance Resilience 6 Mar 2022 · 0 repositories · arXiv:2203.03030
-
Retrieval-Augmented Reinforcement Learning 17 Feb 2022 · 0 repositories · arXiv:2202.08417
-
skrl: Modular and Flexible Library for Reinforcement Learning 8 Feb 2022 · 1 repository · arXiv:2202.03825
-
Generative Adversarial Exploration for Reinforcement Learning 27 Jan 2022 · 0 repositories · arXiv:2201.11685
-
Critic Algorithms using Cooperative Networks 19 Jan 2022 · 0 repositories · arXiv:2201.07839
-
An Improved Reinforcement Learning Algorithm for Learning to Branch 17 Jan 2022 · 0 repositories · arXiv:2201.06213
-
A Resolution Enhancement Plug-in for Deformable Registration of Medical Images 30 Dec 2021 · 0 repositories · arXiv:2112.15180
-
Constraint Sampling Reinforcement Learning: Incorporating Expertise For Faster Learning 30 Dec 2021 · 1 repository · arXiv:2112.15221
-
A Graph Attention Learning Approach to Antenna Tilt Optimization 27 Dec 2021 · 0 repositories · arXiv:2112.14843
-
Intelligent Traffic Light via Policy-based Deep Reinforcement Learning 27 Dec 2021 · 1 repository · arXiv:2112.13817
-
Lane Change Decision-Making through Deep Reinforcement Learning 24 Dec 2021 · 2 repositories · arXiv:2112.14705
-
Local Advantage Networks for Cooperative Multi-Agent Reinforcement Learning 23 Dec 2021 · 0 repositories · arXiv:2112.12458
-
Aerial Base Station Positioning and Power Control for Securing Communications: A Deep Q-Network Approach 21 Dec 2021 · 0 repositories · arXiv:2112.11090
-
A deep reinforcement learning model for predictive maintenance planning of road assets: Integrating LCA and LCCA 20 Dec 2021 · 0 repositories · arXiv:2112.12589
-
Space Non-cooperative Object Active Tracking with Deep Reinforcement Learning 18 Dec 2021 · 1 repository · arXiv:2112.09854
-
Scientific Discovery and the Cost of Measurement -- Balancing Information and Cost in Reinforcement Learning 14 Dec 2021 · 0 repositories · arXiv:2112.07535
-
Faster Deep Reinforcement Learning with Slower Online Network 10 Dec 2021 · 1 repository · arXiv:2112.05848Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
GDI: Rethinking What Makes Reinforcement Learning Different from Supervised Learning 24 Nov 2021 · 0 repositories
-
An Improved Reinforcement Learning Model Based on Sentiment Analysis 19 Nov 2021 · 0 repositories · arXiv:2111.15354
-
Improved Method of Stock Trading under Reinforcement Learning Based on DRQN and Sentiment Indicators ARBR 19 Nov 2021 · 0 repositories · arXiv:2111.15356
-
Improving Experience Replay through Modeling of Similar Transitions' Sets 12 Nov 2021 · 1 repository · arXiv:2111.06907
-
"Good Robot! Now Watch This!": Repurposing Reinforcement Learning for Task-to-Task Transfer 8 Nov 2021 · 1 repository
-
Guiding Multi-Step Rearrangement Tasks with Natural Language Instructions 8 Nov 2021 · 2 repositories
-
Improving RNA Secondary Structure Design using Deep Reinforcement Learning 5 Nov 2021 · 0 repositories · arXiv:2111.04504
-
Online Service Provisioning in NFV-enabled Networks Using Deep Reinforcement Learning 3 Nov 2021 · 0 repositories · arXiv:2111.02209
-
Human-Level Control without Server-Grade Hardware 1 Nov 2021 · 1 repository · arXiv:2111.01264
-
Deep Reinforcement Learning Aided Packet-Routing For Aeronautical Ad-Hoc Networks Formed by Passenger Planes 28 Oct 2021 · 0 repositories · arXiv:2110.15146
-
Accelerating Distributed Deep Reinforcement Learning by In-Network Experience Sampling 26 Oct 2021 · 0 repositories · arXiv:2110.13506
-
Distributional Reinforcement Learning for Multi-Dimensional Reward Functions 26 Oct 2021 · 2 repositories · arXiv:2110.13578
-
Can Q-learning solve Multi Armed Bantids? 21 Oct 2021 · 0 repositories · arXiv:2110.10934
-
More Efficient Exploration with Symbolic Priors on Action Sequence Equivalences 20 Oct 2021 · 0 repositories · arXiv:2110.10632
-
Bridging the Gap between Label- and Reference-based Synthesis in Multi-attribute Image-to-Image Translation 11 Oct 2021 · 1 repository · arXiv:2110.05055
-
Navigation In Urban Environments Amongst Pedestrians Using Multi-Objective Deep Reinforcement Learning 11 Oct 2021 · 0 repositories · arXiv:2110.05205
-
Reinforcement Learning In Two Player Zero Sum Simultaneous Action Games 10 Oct 2021 · 1 repository · arXiv:2110.04835
-
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations 3 Oct 2021 · 0 repositories · arXiv:2110.01101
-
Motion Planning for Autonomous Vehicles in the Presence of Uncertainty Using Reinforcement Learning 1 Oct 2021 · 0 repositories · arXiv:2110.00640
-
Better state exploration using action sequence equivalence 29 Sep 2021 · 0 repositories
-
Bootstrapped Hindsight Experience replay with Counterintuitive Prioritization 29 Sep 2021 · 0 repositories
-
Convergent and Efficient Deep Q Learning Algorithm 29 Sep 2021 · 0 repositories
-
Deep Reinforcement Q-Learning for Intelligent Traffic Signal Control with Partial Detection 29 Sep 2021 · 1 repository · arXiv:2109.14337
-
Disentangling Generalization in Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Explanation-Aware Experience Replay in Rule-Dense Environments 29 Sep 2021 · 1 repository · arXiv:2109.14711
-
HyperDQN: A Randomized Exploration Method for Deep Reinforcement Learning 29 Sep 2021 · 1 repository
-
The guide and the explorer: smart agents for resource-limited iterated batch reinforcement learning 29 Sep 2021 · 0 repositories
-
Fetal oxygen delivery and consumption and blood gases in relation to gestational age 23 Sep 2021 · 0 repositories · arXiv:2109.11616
-
Reinforcement Learning on Encrypted Data 16 Sep 2021 · 0 repositories · arXiv:2109.08236
-
Vision Transformer for Learning Driving Policies in Complex Multi-Agent Environments 14 Sep 2021 · 0 repositories · arXiv:2109.06514
-
Learning cortical representations through perturbed and adversarial dreaming 9 Sep 2021 · 1 repository · arXiv:2109.04261Syntology official (archive's flag): 4 ran · 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples)
-
Temporal Shift Reinforcement Learning 5 Sep 2021 · 1 repository · arXiv:2109.02145
-
Deep Reinforcement Learning at the Edge of the Statistical Precipice 30 Aug 2021 · 3 repositories · arXiv:2108.13264Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 4 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 5 harvested samples)
-
A Microscopic Pandemic Simulator for Pandemic Prediction Using Scalable Million-Agent Reinforcement Learning 14 Aug 2021 · 0 repositories · arXiv:2108.06589
-
DQN Control Solution for KDD Cup 2021 City Brain Challenge 14 Aug 2021 · 1 repository · arXiv:2108.06491
-
Two is a crowd: tracking relations in videos 11 Aug 2021 · 0 repositories · arXiv:2108.05331
-
Modified Double DQN: addressing stability 9 Aug 2021 · 0 repositories · arXiv:2108.04115
-
An Efficient Image-to-Image Translation HourGlass-based Architecture for Object Pushing Policy Learning 2 Aug 2021 · 1 repository · arXiv:2108.01034
-
A DQN-based Approach to Finding Precise Evidences for Fact Verification 1 Aug 2021 · 1 repository
-
REM: Efficient Semi-Automated Real-Time Moderation of Online Forums 1 Aug 2021 · 0 repositories
-
An Improved Algorithm of Robot Path Planning in Complex Environment Based on Double DQN 23 Jul 2021 · 0 repositories · arXiv:2107.11245
-
A Reinforcement Learning Environment for Mathematical Reasoning via Program Synthesis 15 Jul 2021 · 1 repository · arXiv:2107.07373
-
Deep Reinforcement Learning based Dynamic Optimization of Bus Timetable 15 Jul 2021 · 0 repositories · arXiv:2107.07066
-
Minimizing Safety Interference for Safe and Comfortable Automated Driving with Distributional Reinforcement Learning 15 Jul 2021 · 0 repositories · arXiv:2107.07316
-
Improve Agents without Retraining: Parallel Tree Search with Off-Policy Correction 4 Jul 2021 · 1 repository · arXiv:2107.01715
-
AoI Minimization in Energy Harvesting and Spectrum Sharing Enabled 6G Networks 1 Jul 2021 · 0 repositories · arXiv:2107.00340
-
Convergent and Efficient Deep Q Network Algorithm 29 Jun 2021 · 1 repository · arXiv:2106.15419
-
MMD-MIX: Value Function Factorisation with Maximum Mean Discrepancy for Cooperative Multi-Agent Reinforcement Learning 22 Jun 2021 · 0 repositories · arXiv:2106.11652
-
Vision-Language Navigation with Random Environmental Mixup 15 Jun 2021 · 1 repository · arXiv:2106.07876Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
DouZero: Mastering DouDizhu with Self-Play Deep Reinforcement Learning 11 Jun 2021 · 1 repository · arXiv:2106.06135Syntology official (archive's flag): 2 ran · 3 ran (of which 1 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 6 harvested samples) · 1 pointer-only (licence)
-
GDI: Rethinking What Makes Reinforcement Learning Different From Supervised Learning 11 Jun 2021 · 0 repositories · arXiv:2106.06232
-
Reinforced Few-Shot Acquisition Function Learning for Bayesian Optimization 8 Jun 2021 · 0 repositories · arXiv:2106.04335
-
Learning to Optimize Industry-Scale Dynamic Pickup and Delivery Problems 27 May 2021 · 0 repositories · arXiv:2105.12899
-
Improved Exploring Starts by Kernel Density Estimation-Based State-Space Coverage Acceleration in Reinforcement Learning 19 May 2021 · 1 repository · arXiv:2105.08990
-
Reinforcement Learning with Expert Trajectory For Quantitative Trading 9 May 2021 · 0 repositories · arXiv:2105.03844
-
Time-Aware Q-Networks: Resolving Temporal Irregularity for Deep Reinforcement Learning 6 May 2021 · 0 repositories · arXiv:2105.02580
-
Automated scoring of pre-REM sleep in mice with deep learning 5 May 2021 · 1 repository · arXiv:2105.01933
-
Adapting to Reward Progressivity via Spectral Reinforcement Learning 29 Apr 2021 · 1 repository · arXiv:2104.14138Syntology official (archive's flag): 2 ran · 2 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; every one of the 2 samples that ran constructed an object rather than computing a result (of 2 harvested samples) · 2 pointer-only (licence)
-
Emotional Contagion-Aware Deep Reinforcement Learning for Antagonistic Crowd Simulation 29 Apr 2021 · 0 repositories · arXiv:2105.00854
-
Independent Reinforcement Learning for Weakly Cooperative Multiagent Traffic Control Problem 22 Apr 2021 · 1 repository · arXiv:2104.10917
-
A coevolutionary approach to deep multi-agent reinforcement learning 12 Apr 2021 · 1 repository · arXiv:2104.05610
-
Full Gradient DQN Reinforcement Learning: A Provably Convergent Scheme 10 Mar 2021 · 0 repositories · arXiv:2103.05981
-
Increasing Energy Efficiency of Massive-MIMO Network via Base Stations Switching using Reinforcement Learning and Radio Environment Maps 8 Mar 2021 · 0 repositories · arXiv:2103.11891
-
Super-resolving Compressed Images via Parallel and Series Integration of Artifact Reduction and Resolution Enhancement 2 Mar 2021 · 1 repository · arXiv:2103.01698
-
Multi-Agent Path Planning based on MPC and DDPG 26 Feb 2021 · 0 repositories · arXiv:2102.13283
-
Greedy-Step Off-Policy Reinforcement Learning 23 Feb 2021 · 0 repositories · arXiv:2102.11717
-
Stratified Experience Replay: Correcting Multiplicity Bias in Off-Policy Reinforcement Learning 22 Feb 2021 · 0 repositories · arXiv:2102.11319
-
Training a Resilient Q-Network against Observational Interference 18 Feb 2021 · 1 repository · arXiv:2102.09677
-
Adaptive Rational Activations to Boost Deep Reinforcement Learning 18 Feb 2021 · 4 repositories · arXiv:2102.09407Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 6 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 3 pointer-only (licence)