Methods › Reinforcement Learning › Off-Policy TD Control › Q-Learning › Papers, page 6
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,734 · with a code link: 464 · where Syntology ran a sample: 126 (105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (126 of 1,734 tagged: 105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 6 of 18: papers 501 to 600 of 1,734, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Comparing Multiclass Classification Algorithms for Financial Distress Prediction 8 Jul 2023 · 0 repositories · arXiv:2307.03908
-
The Value of Chess Squares 8 Jul 2023 · 0 repositories · arXiv:2307.05330
-
ContainerGym: A Real-World Reinforcement Learning Benchmark for Resource Allocation 6 Jul 2023 · 1 repository · arXiv:2307.02991
-
Offline Reinforcement Learning with Imbalanced Datasets 6 Jul 2023 · 0 repositories · arXiv:2307.02752
-
Interpretable and Secure Trajectory Optimization for UAV-Assisted Communication 5 Jul 2023 · 0 repositories · arXiv:2307.02002
-
Stability of Q-Learning Through Design and Optimism 5 Jul 2023 · 0 repositories · arXiv:2307.02632
-
Achieving Stable Training of Reinforcement Learning Agents in Bimodal Environments through Batch Learning 3 Jul 2023 · 0 repositories · arXiv:2307.00923
-
Is Risk-Sensitive Reinforcement Learning Properly Resolved? 2 Jul 2023 · 0 repositories · arXiv:2307.00547
-
Traceable Group-Wise Self-Optimizing Feature Transformation Learning: A Dual Optimization Perspective 29 Jun 2023 · 1 repository · arXiv:2306.16893
-
Continuous-time q-learning for mean-field control problems 28 Jun 2023 · 0 repositories · arXiv:2306.16208
-
Evaluation of Reinforcement Learning Techniques for Trading on a Diverse Portfolio 28 Jun 2023 · 0 repositories · arXiv:2309.03202
-
Optimizing Credit Limit Adjustments Under Adversarial Goals Using Reinforcement Learning 27 Jun 2023 · 0 repositories · arXiv:2306.15585
-
RansomAI: AI-powered Ransomware for Stealthy Encryption 27 Jun 2023 · 0 repositories · arXiv:2306.15559
-
Decentralized Multi-Robot Formation Control Using Reinforcement Learning 26 Jun 2023 · 0 repositories · arXiv:2306.14489
-
Action Q-Transformer: Visual Explanation in Deep Reinforcement Learning with Encoder-Decoder Model using Action Query 24 Jun 2023 · 0 repositories · arXiv:2306.13879
-
Adaptive Ensemble Q-learning: Minimizing Estimation Bias via Error Feedback 20 Jun 2023 · 0 repositories · arXiv:2306.11918
-
Autonomous Driving with Deep Reinforcement Learning in CARLA Simulation 20 Jun 2023 · 0 repositories · arXiv:2306.11217
-
Vanishing Bias Heuristic-guided Reinforcement Learning Algorithm 17 Jun 2023 · 0 repositories · arXiv:2306.10216
-
Algorithmic Collusion in Auctions: Evidence from Controlled Laboratory Experiments 15 Jun 2023 · 0 repositories · arXiv:2306.09437
-
Joint Path planning and Power Allocation of a Cellular-Connected UAV using Apprenticeship Learning via Deep Inverse Reinforcement Learning 15 Jun 2023 · 1 repository · arXiv:2306.10071
-
Residual Q-Learning: Offline and Online Policy Customization without Value 15 Jun 2023 · 0 repositories · arXiv:2306.09526
-
Privacy Risks in Reinforcement Learning for Household Robots 15 Jun 2023 · 0 repositories · arXiv:2306.09273
-
Model-based versus model-free feeding control and water quality monitoring for fish growth tracking in aquaculture systems 14 Jun 2023 · 0 repositories · arXiv:2306.09915
-
Pruning the Way to Reliable Policies: A Multi-Objective Deep Q-Learning Approach to Critical Care 13 Jun 2023 · 0 repositories · arXiv:2306.08044
-
Approximate information state based convergence analysis of recurrent Q-learning 9 Jun 2023 · 0 repositories · arXiv:2306.05991
-
Finite-Time Analysis of Minimax Q-Learning for Two-Player Zero-Sum Markov Games: Switching System Approach 9 Jun 2023 · 0 repositories · arXiv:2306.05700
-
Quasi-Newton Updating for Large-Scale Distributed Learning 7 Jun 2023 · 0 repositories · arXiv:2306.04111
-
Reinforcement Learning-Based Control of CrazyFlie 2.X Quadrotor 6 Jun 2023 · 0 repositories · arXiv:2306.03951
-
Deep Q-Learning versus Proximal Policy Optimization: Performance Comparison in a Material Sorting Task 2 Jun 2023 · 0 repositories · arXiv:2306.01451
-
IQL-TD-MPC: Implicit Q-Learning for Hierarchical Model Predictive Control 1 Jun 2023 · 0 repositories · arXiv:2306.00867
-
Off-Policy RL Algorithms Can be Sample-Efficient for Continuous Control via Sample Multiple Reuse 29 May 2023 · 1 repository · arXiv:2305.18443
-
VA-learning as a more efficient alternative to Q-learning 29 May 2023 · 0 repositories · arXiv:2305.18161
-
Sample Complexity of Variance-reduced Distributionally Robust Q-learning 28 May 2023 · 0 repositories · arXiv:2305.18420
-
Reinforcement Learning With Reward Machines in Stochastic Games 27 May 2023 · 0 repositories · arXiv:2305.17372
-
Sample Efficient Reinforcement Learning in Mixed Systems through Augmented Samples and Its Applications to Queueing Networks 25 May 2023 · 0 repositories · arXiv:2305.16483
-
RSRM: Reinforcement Symbolic Regression Machine 24 May 2023 · 0 repositories · arXiv:2305.14656
-
When should we prefer Decision Transformers for Offline Reinforcement Learning? 23 May 2023 · 1 repository · arXiv:2305.14550Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 2 pointer-only (licence)
-
Deep Reinforcement Learning-based Multi-objective Path Planning on the Off-road Terrain Environment for Ground Vehicles 23 May 2023 · 0 repositories · arXiv:2305.13783
-
OER: Offline Experience Replay for Continual Offline Reinforcement Learning 23 May 2023 · 0 repositories · arXiv:2305.13804
-
A Framework for Provably Stable and Consistent Training of Deep Feedforward Networks 20 May 2023 · 0 repositories · arXiv:2305.12125
-
Bayesian Risk-Averse Q-Learning with Streaming Observations 18 May 2023 · 0 repositories · arXiv:2305.11300
-
The Blessing of Heterogeneity in Federated Q-Learning: Linear Speedup and Beyond 18 May 2023 · 0 repositories · arXiv:2305.10697
-
How does agency impact human-AI collaborative design space exploration? A case study on ship design with deep generative models 16 May 2023 · 0 repositories · arXiv:2305.10451
-
An Intelligent SDWN Routing Algorithm Based on Network Situational Awareness and Deep Reinforcement Learning 12 May 2023 · 1 repository · arXiv:2305.10441
-
Mastering Percolation-like Games with Deep Learning 12 May 2023 · 1 repository · arXiv:2305.07687
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm 11 May 2023 · 0 repositories · arXiv:2305.06657
-
Extracting Diagnosis Pathways from Electronic Health Records Using Deep Reinforcement Learning 10 May 2023 · 1 repository · arXiv:2305.06295
-
Position Bias Estimation with Item Embedding for Sparse Dataset 10 May 2023 · 0 repositories · arXiv:2305.13931
-
Mixed-Integer Optimal Control via Reinforcement Learning: A Case Study on Hybrid Electric Vehicle Energy Management 2 May 2023 · 1 repository · arXiv:2305.01461
-
BCQQ: Batch-Constraint Quantum Q-Learning with Cyclic Data Re-uploading 27 Apr 2023 · 0 repositories · arXiv:2305.00905
-
Safe Q-learning for continuous-time linear systems 26 Apr 2023 · 0 repositories · arXiv:2304.13573
-
Learned Collusion 25 Apr 2023 · 0 repositories · arXiv:2304.12647
-
IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies 20 Apr 2023 · 1 repository · arXiv:2304.10573Syntology official (archive's flag): 2 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 2 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
Bridging RL Theory and Practice with the Effective Horizon 19 Apr 2023 · 1 repository · arXiv:2304.09853
-
H-TSP: Hierarchically Solving the Large-Scale Travelling Salesman Problem 19 Apr 2023 · 1 repository · arXiv:2304.09395Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
A study on a Q-Learning algorithm application to a manufacturing assembly problem 17 Apr 2023 · 0 repositories · arXiv:2304.08375
-
Collaborative Multi-BS Power Management for Dense Radio Access Network using Deep Reinforcement Learning 17 Apr 2023 · 1 repository · arXiv:2304.07976
-
Exploring the Noise Resilience of Successor Features and Predecessor Features Algorithms in One and Two-Dimensional Environments 14 Apr 2023 · 0 repositories · arXiv:2304.06894
-
Deep reinforcement learning applied to an assembly sequence planning problem with user preferences 13 Apr 2023 · 0 repositories · arXiv:2304.06567
-
Automaton-Guided Curriculum Generation for Reinforcement Learning Agents 11 Apr 2023 · 1 repository · arXiv:2304.05271
-
Reinforcement Learning Based Minimum State-flipped Control for the Reachability of Boolean Control Networks 11 Apr 2023 · 0 repositories · arXiv:2304.04950
-
RELS-DQN: A Robust and Efficient Local Search Framework for Combinatorial Optimization 11 Apr 2023 · 0 repositories · arXiv:2304.06048
-
Generating a Graph Colouring Heuristic with Deep Q-Learning and Graph Neural Networks 8 Apr 2023 · 1 repository · arXiv:2304.04051
-
Deep Reinforcement Learning Based Optimal Infinite-Horizon Control of Probabilistic Boolean Control Networks 7 Apr 2023 · 0 repositories · arXiv:2304.03489
-
Full Gradient Deep Reinforcement Learning for Average-Reward Criterion 7 Apr 2023 · 0 repositories · arXiv:2304.03729
-
Computational role of sleep in memory reorganization 6 Apr 2023 · 0 repositories · arXiv:2304.02873
-
Understanding Reinforcement Learning Algorithms: The Progress from Basic Q-learning to Proximal Policy Optimization 31 Mar 2023 · 0 repositories · arXiv:2304.00026
-
Q-Learning based system for path planning with unmanned aerial vehicles swarms in obstacle environments 30 Mar 2023 · 0 repositories · arXiv:2303.17655
-
Multi-Agent Reinforcement Learning with Action Masking for UAV-enabled Mobile Communications 29 Mar 2023 · 1 repository · arXiv:2303.16737
-
Distributed Multi-Agent Deep Q-Learning for Fast Roaming in IEEE 802.11ax Wi-Fi Systems 25 Mar 2023 · 0 repositories · arXiv:2304.01210
-
Specific investments under negotiated transfer pricing: effects of different surplus sharing parameters on managerial performance: An agent-based simulation with fuzzy Q-learning agents 25 Mar 2023 · 0 repositories · arXiv:2303.14515
-
Robust Path Following on Rivers Using Bootstrapped Reinforcement Learning 24 Mar 2023 · 0 repositories · arXiv:2303.15178
-
Towards Real-World Applications of Personalized Anesthesia Using Policy Constraint Q Learning for Propofol Infusion Control 17 Mar 2023 · 0 repositories · arXiv:2303.10180
-
Self-Inspection Method of Unmanned Aerial Vehicles in Power Plants Using Deep Q-Network Reinforcement Learning 16 Mar 2023 · 0 repositories · arXiv:2303.09013
-
Smoothed Q-learning 15 Mar 2023 · 0 repositories · arXiv:2303.08631
-
Recovering Arrhythmic EEG Transients from Their Stochastic Interference 14 Mar 2023 · 0 repositories · arXiv:2303.07683
-
Digital Twin-Assisted Knowledge Distillation Framework for Heterogeneous Federated Learning 10 Mar 2023 · 0 repositories · arXiv:2303.06155
-
A Framework for History-Aware Hyperparameter Optimisation in Reinforcement Learning 9 Mar 2023 · 0 repositories · arXiv:2303.05186
-
Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning 9 Mar 2023 · 3 repositories · arXiv:2303.05479
-
Learning Strategic Value and Cooperation in Multi-Player Stochastic Games through Side Payments 9 Mar 2023 · 0 repositories · arXiv:2303.05307
-
Environment Transformer and Policy Optimization for Model-Based Offline Reinforcement Learning 7 Mar 2023 · 0 repositories · arXiv:2303.03811
-
Exploration via Epistemic Value Estimation 7 Mar 2023 · 0 repositories · arXiv:2303.04012
-
Double A3C: Deep Reinforcement Learning on OpenAI Gym Games 4 Mar 2023 · 0 repositories · arXiv:2303.02271
-
Wasserstein Actor-Critic: Directed Exploration via Optimism for Continuous-Actions Control 4 Mar 2023 · 0 repositories · arXiv:2303.02378
-
Finite-sample Guarantees for Nash Q-learning with Linear Function Approximation 1 Mar 2023 · 0 repositories · arXiv:2303.00177
-
LS-IQ: Implicit Reward Regularization for Inverse Reinforcement Learning 1 Mar 2023 · 1 repository · arXiv:2303.00599
-
Q-Cogni: An Integrated Causal Reinforcement Learning Framework 26 Feb 2023 · 0 repositories · arXiv:2302.13240
-
On Bellman's principle of optimality and Reinforcement learning for safety-constrained Markov decision process 25 Feb 2023 · 0 repositories · arXiv:2302.13152
-
Gauss-Newton Temporal Difference Learning with Nonlinear Function Approximation 25 Feb 2023 · 0 repositories · arXiv:2302.13087
-
Kernel-Based Distributed Q-Learning: A Scalable Reinforcement Learning Approach for Dynamic Treatment Regimes 21 Feb 2023 · 0 repositories · arXiv:2302.10434
-
Learning to Play Text-based Adventure Games with Maximum Entropy Reinforcement Learning 21 Feb 2023 · 1 repository · arXiv:2302.10720
-
Robust Auto-landing Control of an agile Regional Jet Using Fuzzy Q-learning 21 Feb 2023 · 0 repositories · arXiv:2302.10997
-
Forecasting and stabilizing chaotic regimes in two macroeconomic models via artificial intelligence technologies and control methods 20 Feb 2023 · 0 repositories · arXiv:2302.12019
-
Deep Offline Reinforcement Learning for Real-world Treatment Optimization Applications 15 Feb 2023 · 0 repositories · arXiv:2302.07549
-
Online Statistical Inference for Nonlinear Stochastic Approximation with Markovian Data 15 Feb 2023 · 0 repositories · arXiv:2302.07690
-
A Lifetime Extended Energy Management Strategy for Fuel Cell Hybrid Electric Vehicles via Self-Learning Fuzzy Reinforcement Learning 13 Feb 2023 · 0 repositories · arXiv:2302.06236
-
Computation Offloading for Uncertain Marine Tasks by Cooperation of UAVs and Vessels 13 Feb 2023 · 0 repositories · arXiv:2302.06055
-
Differentially Private Deep Q-Learning for Pattern Privacy Preservation in MEC Offloading 9 Feb 2023 · 0 repositories · arXiv:2302.04608
-
Catch Me If You Can: Improving Adversaries in Cyber-Security With Q-Learning Algorithms 7 Feb 2023 · 0 repositories · arXiv:2302.03768
-
Ensemble Value Functions for Efficient Exploration in Multi-Agent Reinforcement Learning 7 Feb 2023 · 0 repositories · arXiv:2302.03439