Methods › Reinforcement Learning › Off-Policy TD Control › Q-Learning › Papers, page 9
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,734 · with a code link: 464 · where Syntology ran a sample: 126 (105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (126 of 1,734 tagged: 105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 9 of 18: papers 801 to 900 of 1,734, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Multiple Correlated Jammers Nullification using LSTM-based Deep Dueling Neural Network 8 Feb 2022 · 0 repositories · arXiv:2202.03600
-
skrl: Modular and Flexible Library for Reinforcement Learning 8 Feb 2022 · 1 repository · arXiv:2202.03825
-
Transfer Reinforcement Learning for Differing Action Spaces via Q-Network Representations 5 Feb 2022 · 1 repository · arXiv:2202.02442
-
Generative Adversarial Exploration for Reinforcement Learning 27 Jan 2022 · 0 repositories · arXiv:2201.11685
-
Deep Q-learning: a robust control approach 21 Jan 2022 · 1 repository · arXiv:2201.08610
-
Critic Algorithms using Cooperative Networks 19 Jan 2022 · 0 repositories · arXiv:2201.07839
-
An Improved Reinforcement Learning Algorithm for Learning to Branch 17 Jan 2022 · 0 repositories · arXiv:2201.06213
-
A Family of Cognitively Realistic Parsing Environments for Deep Reinforcement Learning 16 Jan 2022 · 0 repositories
-
Criticality-Based Varying Step-Number Algorithm for Reinforcement Learning 13 Jan 2022 · 0 repositories · arXiv:2201.05034
-
Age-of-information minimization via opportunistic sampling by an energy harvesting source 8 Jan 2022 · 0 repositories · arXiv:2201.02787
-
Sales Time Series Analytics Using Deep Q-Learning 6 Jan 2022 · 0 repositories · arXiv:2201.02058
-
Reinforcement Learning for Task Specifications with Action-Constraints 2 Jan 2022 · 0 repositories · arXiv:2201.00286
-
Operator Deep Q-Learning: Zero-Shot Reward Transferring in Reinforcement Learning 1 Jan 2022 · 0 repositories · arXiv:2201.00236
-
A Resolution Enhancement Plug-in for Deformable Registration of Medical Images 30 Dec 2021 · 0 repositories · arXiv:2112.15180
-
Constraint Sampling Reinforcement Learning: Incorporating Expertise For Faster Learning 30 Dec 2021 · 1 repository · arXiv:2112.15221
-
A Statistical Analysis of Polyak-Ruppert Averaged Q-learning 29 Dec 2021 · 1 repository · arXiv:2112.14582Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Graph Attention Learning Approach to Antenna Tilt Optimization 27 Dec 2021 · 0 repositories · arXiv:2112.14843
-
Intelligent Traffic Light via Policy-based Deep Reinforcement Learning 27 Dec 2021 · 1 repository · arXiv:2112.13817
-
Lane Change Decision-Making through Deep Reinforcement Learning 24 Dec 2021 · 2 repositories · arXiv:2112.14705
-
Local Advantage Networks for Cooperative Multi-Agent Reinforcement Learning 23 Dec 2021 · 0 repositories · arXiv:2112.12458
-
Safety and Liveness Guarantees through Reach-Avoid Reinforcement Learning 23 Dec 2021 · 1 repository · arXiv:2112.12288
-
Aerial Base Station Positioning and Power Control for Securing Communications: A Deep Q-Network Approach 21 Dec 2021 · 0 repositories · arXiv:2112.11090
-
A deep reinforcement learning model for predictive maintenance planning of road assets: Integrating LCA and LCCA 20 Dec 2021 · 0 repositories · arXiv:2112.12589
-
Space Non-cooperative Object Active Tracking with Deep Reinforcement Learning 18 Dec 2021 · 1 repository · arXiv:2112.09854
-
Finite-Sample Analysis of Decentralized Q-Learning for Stochastic Games 15 Dec 2021 · 0 repositories · arXiv:2112.07859
-
Scientific Discovery and the Cost of Measurement -- Balancing Information and Cost in Reinforcement Learning 14 Dec 2021 · 0 repositories · arXiv:2112.07535
-
Teaching a Robot to Walk Using Reinforcement Learning 13 Dec 2021 · 0 repositories · arXiv:2112.07031
-
Control-Tutored Reinforcement Learning: Towards the Integration of Data-Driven and Model-Based Control 11 Dec 2021 · 0 repositories · arXiv:2112.06018
-
Faster Deep Reinforcement Learning with Slower Online Network 10 Dec 2021 · 1 repository · arXiv:2112.05848Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
Quantum Architecture Search via Continual Reinforcement Learning 10 Dec 2021 · 0 repositories · arXiv:2112.05779
-
Application of Deep Reinforcement Learning to Payment Fraud 8 Dec 2021 · 0 repositories · arXiv:2112.04236
-
Convergence Results For Q-Learning With Experience Replay 8 Dec 2021 · 0 repositories · arXiv:2112.04213
-
Replay For Safety 8 Dec 2021 · 0 repositories · arXiv:2112.04229
-
A Risk-Averse Preview-based Q-Learning Algorithm: Application to Highway Driving of Autonomous Vehicles 6 Dec 2021 · 0 repositories · arXiv:2112.03232
-
Faster Non-asymptotic Convergence for Double Q-learning 1 Dec 2021 · 0 repositories
-
Solving reward-collecting problems with UAVs: a comparison of online optimization and Q-learning 30 Nov 2021 · 1 repository · arXiv:2112.00141
-
DeepCQ+: Robust and Scalable Routing with Multi-Agent Deep Reinforcement Learning for Highly Dynamic Networks 29 Nov 2021 · 0 repositories · arXiv:2111.15013
-
Final Adaptation Reinforcement Learning for N-Player Games 29 Nov 2021 · 0 repositories · arXiv:2111.14375
-
Count-Based Temperature Scheduling for Maximum Entropy Reinforcement Learning 28 Nov 2021 · 0 repositories · arXiv:2111.14204
-
Deep Q-Learning based Reinforcement Learning Approach for Network Intrusion Detection 27 Nov 2021 · 1 repository · arXiv:2111.13978
-
GDI: Rethinking What Makes Reinforcement Learning Different from Supervised Learning 24 Nov 2021 · 0 repositories
-
Multicrew Scheduling and Routing in Road Network Restoration Based on Deep Q-learning 24 Nov 2021 · 0 repositories
-
Reversible Action Design for Combinatorial Optimization with ReinforcementLearning 24 Nov 2021 · 0 repositories
-
The Impact of Data Distribution on Q-learning with Function Approximation 23 Nov 2021 · 1 repository · arXiv:2111.11758
-
Multi-agent Bayesian Deep Reinforcement Learning for Microgrid Energy Management under Communication Failures 22 Nov 2021 · 0 repositories · arXiv:2111.11868
-
An Improved Reinforcement Learning Model Based on Sentiment Analysis 19 Nov 2021 · 0 repositories · arXiv:2111.15354
-
Improved Method of Stock Trading under Reinforcement Learning Based on DRQN and Sentiment Indicators ARBR 19 Nov 2021 · 0 repositories · arXiv:2111.15356
-
Aggressive Q-Learning with Ensembles: Achieving Both High Sample Efficiency and High Asymptotic Performance 17 Nov 2021 · 0 repositories · arXiv:2111.09159
-
Consecutive Task-oriented Dialog Policy Learning 16 Nov 2021 · 0 repositories
-
Where to Look: A Unified Attention Model for Visual Recognition with Reinforcement Learning 13 Nov 2021 · 0 repositories · arXiv:2111.07169
-
Improving Experience Replay through Modeling of Similar Transitions' Sets 12 Nov 2021 · 1 repository · arXiv:2111.06907
-
Q-Learning for MDPs with General Spaces: Convergence and Near Optimality via Quantization under Weak Continuity 12 Nov 2021 · 0 repositories · arXiv:2111.06781
-
"Good Robot! Now Watch This!": Repurposing Reinforcement Learning for Task-to-Task Transfer 8 Nov 2021 · 1 repository
-
Guiding Multi-Step Rearrangement Tasks with Natural Language Instructions 8 Nov 2021 · 2 repositories
-
On Assessing The Safety of Reinforcement Learning algorithms Using Formal Methods 8 Nov 2021 · 0 repositories · arXiv:2111.04865
-
Improving RNA Secondary Structure Design using Deep Reinforcement Learning 5 Nov 2021 · 0 repositories · arXiv:2111.04504
-
Supervised Advantage Actor-Critic for Recommender Systems 5 Nov 2021 · 0 repositories · arXiv:2111.03474
-
Balanced Q-learning: Combining the Influence of Optimistic and Pessimistic Targets 3 Nov 2021 · 0 repositories · arXiv:2111.02787
-
Online Service Provisioning in NFV-enabled Networks Using Deep Reinforcement Learning 3 Nov 2021 · 0 repositories · arXiv:2111.02209
-
Koopman Q-learning: Offline Reinforcement Learning via Symmetries of Dynamics 2 Nov 2021 · 0 repositories · arXiv:2111.01365
-
Human-Level Control without Server-Grade Hardware 1 Nov 2021 · 1 repository · arXiv:2111.01264
-
Decentralized Multi-Agent Reinforcement Learning: An Off-Policy Method 31 Oct 2021 · 0 repositories · arXiv:2111.00438
-
Throughput and Latency in the Distributed Q-Learning Random Access mMTC Networks 30 Oct 2021 · 0 repositories · arXiv:2111.00299
-
Learning to Communicate with Reinforcement Learning for an Adaptive Traffic Control System 29 Oct 2021 · 0 repositories · arXiv:2110.15779
-
Location-routing Optimisation for Urban Logistics Using Mobile Parcel Locker Based on Hybrid Q-Learning Algorithm 29 Oct 2021 · 0 repositories · arXiv:2110.15485
-
Deep Reinforcement Learning Aided Packet-Routing For Aeronautical Ad-Hoc Networks Formed by Passenger Planes 28 Oct 2021 · 0 repositories · arXiv:2110.15146
-
Temporal-Difference Value Estimation via Uncertainty-Guided Soft Updates 28 Oct 2021 · 0 repositories · arXiv:2110.14818
-
Finite Horizon Q-learning: Stability, Convergence, Simulations and an application on Smart Grids 27 Oct 2021 · 0 repositories · arXiv:2110.15093
-
V-Learning -- A Simple, Efficient, Decentralized Algorithm for Multiagent RL 27 Oct 2021 · 0 repositories · arXiv:2110.14555
-
Accelerating Distributed Deep Reinforcement Learning by In-Network Experience Sampling 26 Oct 2021 · 0 repositories · arXiv:2110.13506
-
Distributional Reinforcement Learning for Multi-Dimensional Reward Functions 26 Oct 2021 · 2 repositories · arXiv:2110.13578
-
Multi-Agent Advisor Q-Learning 26 Oct 2021 · 1 repository · arXiv:2111.00345
-
Can Q-learning solve Multi Armed Bantids? 21 Oct 2021 · 0 repositories · arXiv:2110.10934
-
More Efficient Exploration with Symbolic Priors on Action Sequence Equivalences 20 Oct 2021 · 0 repositories · arXiv:2110.10632
-
Playing 2048 With Reinforcement Learning 20 Oct 2021 · 1 repository · arXiv:2110.10374
-
A Q-Learning-based Approach for Distributed Beam Scheduling in mmWave Networks 17 Oct 2021 · 0 repositories · arXiv:2110.08704
-
Online Target Q-learning with Reverse Experience Replay: Efficiently finding the Optimal Policy for Linear MDPs 16 Oct 2021 · 0 repositories · arXiv:2110.08440
-
Value Penalized Q-Learning for Recommender Systems 15 Oct 2021 · 0 repositories · arXiv:2110.07923
-
On Improving Model-Free Algorithms for Decentralized Multi-Agent Reinforcement Learning 12 Oct 2021 · 0 repositories · arXiv:2110.05707
-
Fast Block Linear System Solver Using Q-Learning Schduling for Unified Dynamic Power System Simulations 12 Oct 2021 · 0 repositories · arXiv:2110.05843
-
Bridging the Gap between Label- and Reference-based Synthesis in Multi-attribute Image-to-Image Translation 11 Oct 2021 · 1 repository · arXiv:2110.05055
-
Navigation In Urban Environments Amongst Pedestrians Using Multi-Objective Deep Reinforcement Learning 11 Oct 2021 · 0 repositories · arXiv:2110.05205
-
A Deep Learning Inference Scheme Based on Pipelined Matrix Multiplication Acceleration Design and Non-uniform Quantization 10 Oct 2021 · 0 repositories · arXiv:2110.04861
-
Reinforcement Learning In Two Player Zero Sum Simultaneous Action Games 10 Oct 2021 · 1 repository · arXiv:2110.04835
-
Breaking the Sample Complexity Barrier to Regret-Optimal Model-Free Reinforcement Learning 9 Oct 2021 · 0 repositories · arXiv:2110.04645
-
Training Transition Policies via Distribution Matching for Complex Tasks 8 Oct 2021 · 1 repository · arXiv:2110.04357
-
Deep reinforcement learning for guidewire navigation in coronary artery phantom 5 Oct 2021 · 0 repositories · arXiv:2110.01840
-
Dropout Q-Functions for Doubly Efficient Reinforcement Learning 5 Oct 2021 · 2 repositories · arXiv:2110.02034Syntology official (archive's flag): 1 ran · 6 ran (of which 3 constructed an object rather than computing a result; 6 with no instrument failure: 2 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 8 harvested samples)
-
Uncertainty-Based Offline Reinforcement Learning with Diversified Q-Ensemble 4 Oct 2021 · 5 repositories · arXiv:2110.01548Syntology 13 ran (of which 11 constructed an object rather than computing a result; 12 with no instrument failure: 0 honoured, 0 violated, 12 with no contract checked; 1 where Syntology's instrument failed) · 8 unverified (of 21 harvested samples) · 6 pointer-only (licence)
-
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations 3 Oct 2021 · 0 repositories · arXiv:2110.01101
-
Cellular traffic offloading via Opportunistic Networking with Reinforcement Learning 1 Oct 2021 · 0 repositories · arXiv:2110.00397
-
Motion Planning for Autonomous Vehicles in the Presence of Uncertainty Using Reinforcement Learning 1 Oct 2021 · 0 repositories · arXiv:2110.00640
-
Learning the Markov Decision Process in the Sparse Gaussian Elimination 30 Sep 2021 · 1 repository · arXiv:2109.14929
-
Adaptive Q-learning for Interaction-Limited Reinforcement Learning 29 Sep 2021 · 0 repositories
-
An Attempt to Model Human Trust with Reinforcement Learning 29 Sep 2021 · 0 repositories
-
Better state exploration using action sequence equivalence 29 Sep 2021 · 0 repositories
-
Bootstrapped Hindsight Experience replay with Counterintuitive Prioritization 29 Sep 2021 · 0 repositories
-
Convergent and Efficient Deep Q Learning Algorithm 29 Sep 2021 · 0 repositories
-
Decentralized Cooperative Multi-Agent Reinforcement Learning with Exploration 29 Sep 2021 · 0 repositories
-
Deep Reinforcement Q-Learning for Intelligent Traffic Signal Control with Partial Detection 29 Sep 2021 · 1 repository · arXiv:2109.14337