Methods › Reinforcement Learning › Q-Learning Networks › DQN › Papers, page 2
Deep Q-Network
DQN
Papers archive 2025-07-28
archive papers tagged: 519 · with a code link: 173 · where Syntology ran a sample: 47 (36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (47 of 519 tagged: 36 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 2 of 6: papers 101 to 200 of 519, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Wake-Sleep Consolidated Learning 6 Dec 2023 · 0 repositories · arXiv:2401.08623
-
Lights out: training RL agents robust to temporary blindness 5 Dec 2023 · 0 repositories · arXiv:2312.02665
-
AdsorbRL: Deep Multi-Objective Reinforcement Learning for Inverse Catalysts Design 4 Dec 2023 · 1 repository · arXiv:2312.02308Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Self-Driving Telescopes: Autonomous Scheduling of Astronomical Observation Campaigns with Offline Reinforcement Learning 29 Nov 2023 · 0 repositories · arXiv:2311.18094
-
Reinforcement Learning for Wildfire Mitigation in Simulated Disaster Environments 27 Nov 2023 · 1 repository · arXiv:2311.15925Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Two-compartment neuronal spiking model expressing brain-state specific apical-amplification, -isolation and -drive regimes 10 Nov 2023 · 0 repositories · arXiv:2311.06074
-
Advancing Algorithmic Trading: A Multi-Technique Enhancement of Deep Q-Network Models 9 Nov 2023 · 0 repositories · arXiv:2311.05743
-
Selectively Sharing Experiences Improves Multi-Agent Reinforcement Learning 1 Nov 2023 · 1 repository · arXiv:2311.00865Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning 29 Oct 2023 · 0 repositories · arXiv:2310.19137
-
Weakly Coupled Deep Q-Networks 28 Oct 2023 · 0 repositories · arXiv:2310.18803Syntology 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
On the Convergence and Sample Complexity Analysis of Deep Q-Networks with ε-Greedy Exploration 24 Oct 2023 · 0 repositories · arXiv:2310.16173
-
Reinforcement learning based local path planning for mobile robot 24 Oct 2023 · 0 repositories · arXiv:2403.12463
-
Deep Reinforcement Learning-based Intelligent Traffic Signal Controls with Optimized CO2 emissions 19 Oct 2023 · 1 repository · arXiv:2310.13129
-
Automatic Music Playlist Generation via Simulation-based Reinforcement Learning 13 Oct 2023 · 0 repositories · arXiv:2310.09123
-
Optimal Scheduling of Electric Vehicle Charging with Deep Reinforcement Learning considering End Users Flexibility 13 Oct 2023 · 0 repositories · arXiv:2310.09040
-
Learning RL-Policies for Joint Beamforming Without Exploration: A Batch Constrained Off-Policy Approach 12 Oct 2023 · 1 repository · arXiv:2310.08660
-
Optimal Sequential Decision-Making in Geosteering: A Reinforcement Learning Approach 7 Oct 2023 · 0 repositories · arXiv:2310.04772
-
PGDQN: Preference-Guided Deep Q-Network 3 Oct 2023 · 1 repository
-
SI-SD: Sleep Interpreter through awake-guided cross-subject Semantic Decoding 28 Sep 2023 · 0 repositories · arXiv:2309.16457
-
Deep Reinforcement Learning for the Heat Transfer Control of Pulsating Impinging Jets 25 Sep 2023 · 0 repositories · arXiv:2309.13955
-
Enhancing data efficiency in reinforcement learning: a novel imagination mechanism based on mesh information propagation 25 Sep 2023 · 2 repositories · arXiv:2309.14243
-
Enhancing Healthcare with EOG: A Novel Approach to Sleep Stage Classification 25 Sep 2023 · 0 repositories · arXiv:2310.03757
-
Implicit Sensing in Traffic Optimization: Advanced Deep Reinforcement Learning Techniques 25 Sep 2023 · 0 repositories · arXiv:2309.14395
-
On the Convergence and Sample Complexity Analysis of Deep Q-Networks with ϵ-Greedy Exploration 21 Sep 2023 · 0 repositories
-
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions 20 Sep 2023 · 0 repositories · arXiv:2309.10980
-
Coalescent processes emerging from large deviations 28 Aug 2023 · 0 repositories · arXiv:2308.14715
-
Prompt to Transfer: Sim-to-Real Transfer for Traffic Signal Control with Prompt Learning 28 Aug 2023 · 1 repository · arXiv:2308.14284Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Graph Neural Network-Based QUBO-Formulated Hamiltonian-Inspired Loss Function for Combinatorial Optimization using Reinforcement Learning 27 Aug 2023 · 0 repositories · arXiv:2308.13978
-
Integrating Renewable Energy in Agriculture: A Deep Reinforcement Learning-based Approach 16 Aug 2023 · 0 repositories · arXiv:2308.08611
-
Bag of Policies for Distributional Deep Exploration 3 Aug 2023 · 0 repositories · arXiv:2308.01759
-
Caching-at-STARS: the Next Generation Edge Caching 1 Aug 2023 · 0 repositories · arXiv:2308.00562
-
Pixel to policy: DQN Encoders for within & cross-game reinforcement learning 1 Aug 2023 · 0 repositories · arXiv:2308.00318
-
ETHER: Aligning Emergent Communication for Hindsight Experience Replay 28 Jul 2023 · 0 repositories · arXiv:2307.15494
-
Goal-Conditioned Reinforcement Learning with Disentanglement-based Reachability Planning 20 Jul 2023 · 0 repositories · arXiv:2307.10846
-
Distributed 3D-Beam Reforming for Hovering-Tolerant UAVs Communication over Coexistence: A Deep-Q Learning for Intelligent Space-Air-Ground Integrated Networks 18 Jul 2023 · 0 repositories · arXiv:2307.09325
-
Measuring and Mitigating Interference in Reinforcement Learning 10 Jul 2023 · 0 repositories · arXiv:2307.04887
-
Probabilistic Counterexample Guidance for Safer Reinforcement Learning (Extended Version) 10 Jul 2023 · 1 repository · arXiv:2307.04927
-
Investigating the Edge of Stability Phenomenon in Reinforcement Learning 9 Jul 2023 · 0 repositories · arXiv:2307.04210
-
Comparing Multiclass Classification Algorithms for Financial Distress Prediction 8 Jul 2023 · 0 repositories · arXiv:2307.03908
-
ContainerGym: A Real-World Reinforcement Learning Benchmark for Resource Allocation 6 Jul 2023 · 1 repository · arXiv:2307.02991
-
Interpretable and Secure Trajectory Optimization for UAV-Assisted Communication 5 Jul 2023 · 0 repositories · arXiv:2307.02002
-
Vanishing Bias Heuristic-guided Reinforcement Learning Algorithm 17 Jun 2023 · 0 repositories · arXiv:2306.10216
-
Quasi-Newton Updating for Large-Scale Distributed Learning 7 Jun 2023 · 0 repositories · arXiv:2306.04111
-
Deep Q-Learning versus Proximal Policy Optimization: Performance Comparison in a Material Sorting Task 2 Jun 2023 · 0 repositories · arXiv:2306.01451
-
VA-learning as a more efficient alternative to Q-learning 29 May 2023 · 0 repositories · arXiv:2305.18161
-
Deep Reinforcement Learning-based Multi-objective Path Planning on the Off-road Terrain Environment for Ground Vehicles 23 May 2023 · 0 repositories · arXiv:2305.13783
-
How does agency impact human-AI collaborative design space exploration? A case study on ship design with deep generative models 16 May 2023 · 0 repositories · arXiv:2305.10451
-
An Intelligent SDWN Routing Algorithm Based on Network Situational Awareness and Deep Reinforcement Learning 12 May 2023 · 1 repository · arXiv:2305.10441
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm 11 May 2023 · 0 repositories · arXiv:2305.06657
-
Extracting Diagnosis Pathways from Electronic Health Records Using Deep Reinforcement Learning 10 May 2023 · 1 repository · arXiv:2305.06295
-
Position Bias Estimation with Item Embedding for Sparse Dataset 10 May 2023 · 0 repositories · arXiv:2305.13931
-
Bridging RL Theory and Practice with the Effective Horizon 19 Apr 2023 · 1 repository · arXiv:2304.09853
-
Collaborative Multi-BS Power Management for Dense Radio Access Network using Deep Reinforcement Learning 17 Apr 2023 · 1 repository · arXiv:2304.07976
-
RELS-DQN: A Robust and Efficient Local Search Framework for Combinatorial Optimization 11 Apr 2023 · 0 repositories · arXiv:2304.06048
-
Full Gradient Deep Reinforcement Learning for Average-Reward Criterion 7 Apr 2023 · 0 repositories · arXiv:2304.03729
-
Computational role of sleep in memory reorganization 6 Apr 2023 · 0 repositories · arXiv:2304.02873
-
Multi-Agent Reinforcement Learning with Action Masking for UAV-enabled Mobile Communications 29 Mar 2023 · 1 repository · arXiv:2303.16737
-
Self-Inspection Method of Unmanned Aerial Vehicles in Power Plants Using Deep Q-Network Reinforcement Learning 16 Mar 2023 · 0 repositories · arXiv:2303.09013
-
Recovering Arrhythmic EEG Transients from Their Stochastic Interference 14 Mar 2023 · 0 repositories · arXiv:2303.07683
-
A Framework for History-Aware Hyperparameter Optimisation in Reinforcement Learning 9 Mar 2023 · 0 repositories · arXiv:2303.05186
-
Gauss-Newton Temporal Difference Learning with Nonlinear Function Approximation 25 Feb 2023 · 0 repositories · arXiv:2302.13087
-
Ensemble Value Functions for Efficient Exploration in Multi-Agent Reinforcement Learning 7 Feb 2023 · 0 repositories · arXiv:2302.03439
-
Deep Reinforcement Learning for Traffic Light Control in Intelligent Transportation Systems 4 Feb 2023 · 0 repositories · arXiv:2302.03669
-
Accelerating Policy Gradient by Estimating Value Function from Prior Computation in Deep Reinforcement Learning 2 Feb 2023 · 0 repositories · arXiv:2302.01399
-
Sample Efficient Deep Reinforcement Learning via Local Planning 29 Jan 2023 · 0 repositories · arXiv:2301.12579
-
schlably: A Python Framework for Deep Reinforcement Learning Based Scheduling Experiments 10 Jan 2023 · 1 repository · arXiv:2301.04182
-
XDQN: Inherently Interpretable DQN through Mimicking 8 Jan 2023 · 0 repositories · arXiv:2301.03043
-
Hierarchical Deep Reinforcement Learning for Age-of-Information Minimization in IRS-aided and Wireless-powered Wireless Networks 27 Dec 2022 · 0 repositories · arXiv:2212.13390
-
Control of Continuous Quantum Systems with Many Degrees of Freedom based on Convergent Reinforcement Learning 21 Dec 2022 · 1 repository · arXiv:2212.10705
-
Neighboring state-based RL Exploration 21 Dec 2022 · 0 repositories · arXiv:2212.10712
-
Airfoil Shape Optimization using Deep Q-Network 29 Nov 2022 · 0 repositories · arXiv:2211.17189
-
Applying Deep Reinforcement Learning to the HP Model for Protein Structure Prediction 27 Nov 2022 · 1 repository · arXiv:2211.14939Syntology official (archive's flag): 11 ran · 11 ran (of which 0 constructed an object rather than computing a result; 11 with no instrument failure: 0 honoured, 0 violated, 11 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 13 harvested samples)
-
Simultaneously Updating All Persistence Values in Reinforcement Learning 21 Nov 2022 · 0 repositories · arXiv:2211.11620
-
Solar Power driven EV Charging Optimization with Deep Reinforcement Learning 17 Nov 2022 · 0 repositories · arXiv:2211.09479
-
NREM and REM: cognitive and energetic gains in thalamo-cortical sleeping and awake spiking model 13 Nov 2022 · 0 repositories · arXiv:2211.06889
-
Deep W-Networks: Solving Multi-Objective Optimisation Problems With Deep Reinforcement Learning 9 Nov 2022 · 1 repository · arXiv:2211.04813
-
Deep Reinforcement Learning for Power Control in Next-Generation WiFi Network Systems 2 Nov 2022 · 0 repositories · arXiv:2211.01107
-
Representation Learning for General-sum Low-rank Markov Games 30 Oct 2022 · 0 repositories · arXiv:2210.16976
-
DeFIX: Detecting and Fixing Failure Scenarios with Reinforcement Learning in Imitation Learning Based Autonomous Driving 29 Oct 2022 · 2 repositories · arXiv:2210.16567
-
Elastic Step DQN: A novel multi-step algorithm to alleviate overestimation in Deep QNetworks 7 Oct 2022 · 0 repositories · arXiv:2210.03325
-
Exploration Policies for On-the-Fly Controller Synthesis: A Reinforcement Learning Approach 7 Oct 2022 · 1 repository · arXiv:2210.05393
-
M²DQN: A Robust Method for Accelerating Deep Q-learning Network 16 Sep 2022 · 1 repository · arXiv:2209.07809
-
Reducing Variance in Temporal-Difference Value Estimation via Ensemble of Deep Networks 16 Sep 2022 · 1 repository · arXiv:2209.07670
-
Pathfinding in Random Partially Observable Environments with Vision-Informed Deep Reinforcement Learning 11 Sep 2022 · 0 repositories · arXiv:2209.04801
-
Continual learning benefits from multiple sleep mechanisms: NREM, REM, and Synaptic Downscaling 9 Sep 2022 · 0 repositories · arXiv:2209.05245
-
Distilling Deep RL Models Into Interpretable Neuro-Fuzzy Systems 7 Sep 2022 · 0 repositories · arXiv:2209.03357
-
Disentangled Modeling of Domain and Relevance for Adaptable Dense Retrieval 11 Aug 2022 · 1 repository · arXiv:2208.05753
-
Prediction-based Hybrid Slicing Framework for Service Level Agreement Guarantee in Mobility Scenarios: A Deep Learning Approach 6 Aug 2022 · 0 repositories · arXiv:2208.03460
-
A Maintenance Planning Framework using Online and Offline Deep Reinforcement Learning 1 Aug 2022 · 0 repositories · arXiv:2208.00808
-
A Deep Reinforcement Learning Approach for Finding Non-Exploitable Strategies in Two-Player Atari Games 18 Jul 2022 · 2 repositories · arXiv:2207.08894Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples)
-
Boolean Decision Rules for Reinforcement Learning Policy Summarisation 18 Jul 2022 · 0 repositories · arXiv:2207.08651
-
Deep Reinforcement Learning with Swin Transformers 30 Jun 2022 · 1 repository · arXiv:2206.15269
-
DNA: Proximal Policy Optimization with a Dual Network Architecture 20 Jun 2022 · 1 repository · arXiv:2206.10027
-
Sampling Efficient Deep Reinforcement Learning through Preference-Guided Stochastic Exploration 20 Jun 2022 · 1 repository · arXiv:2206.09627
-
Goal-Space Planning with Subgoal Models 6 Jun 2022 · 0 repositories · arXiv:2206.02902
-
The Phenomenon of Policy Churn 1 Jun 2022 · 0 repositories · arXiv:2206.00730
-
Efficient Reward Poisoning Attacks on Online Deep Reinforcement Learning 30 May 2022 · 1 repository · arXiv:2205.14842Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Improving Bidding and Playing Strategies in the Trick-Taking game Wizard using Deep Q-Networks 27 May 2022 · 0 repositories · arXiv:2205.13834
-
Multiple Domain Cyberspace Attack and Defense Game Based on Reward Randomization Reinforcement Learning 23 May 2022 · 0 repositories · arXiv:2205.10990
-
Long Run Incremental Cost (LRIC) Distribution Network Pricing in UK, advising China's Distribution Network 20 May 2022 · 0 repositories · arXiv:2205.09946