Methods › Reinforcement Learning › Off-Policy TD Control › Q-Learning › Papers, page 4
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,734 · with a code link: 464 · where Syntology ran a sample: 126 (105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (126 of 1,734 tagged: 105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 4 of 18: papers 301 to 400 of 1,734, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Data-Incremental Continual Offline Reinforcement Learning 19 Apr 2024 · 0 repositories · arXiv:2404.12639
-
From r to Q^*: Your Language Model is Secretly a Q-Function 18 Apr 2024 · 0 repositories · arXiv:2404.12358
-
Empowering Embodied Visual Tracking with Visual Foundation Models and Offline RL 15 Apr 2024 · 0 repositories · arXiv:2404.09857
-
Advancing Forest Fire Prevention: Deep Reinforcement Learning for Effective Firebreak Placement 12 Apr 2024 · 0 repositories · arXiv:2404.08523
-
Traffic Signal Control and Speed Offset Coordination Using Q-Learning for Arterial Road Networks 9 Apr 2024 · 0 repositories · arXiv:2404.06382
-
Deep Reinforcement Learning Control for Disturbance Rejection in a Nonlinear Dynamic System with Parametric Uncertainty 6 Apr 2024 · 0 repositories · arXiv:2404.04699
-
Growing Q-Networks: Solving Continuous Control Tasks with Adaptive Control Resolution 5 Apr 2024 · 0 repositories · arXiv:2404.04253
-
Superior Genetic Algorithms for the Target Set Selection Problem Based on Power-Law Parameter Choices and Simple Greedy Heuristics 5 Apr 2024 · 1 repository · arXiv:2404.04018
-
Laser Learning Environment: A new environment for coordination-critical multi-agent tasks 4 Apr 2024 · 1 repository · arXiv:2404.03596
-
K-percent Evaluation for Lifelong RL 2 Apr 2024 · 0 repositories · arXiv:2404.02113
-
Utilizing Maximum Mean Discrepancy Barycenter for Propagating the Uncertainty of Value Functions in Reinforcement Learning 31 Mar 2024 · 0 repositories · arXiv:2404.00686
-
EnCoMP: Enhanced Covert Maneuver Planning with Adaptive Threat-Aware Visibility Estimation using Offline Reinforcement Learning 29 Mar 2024 · 0 repositories · arXiv:2403.20016
-
From Two-Dimensional to Three-Dimensional Environment with Q-Learning: Modeling Autonomous Navigation with Reinforcement Learning and no Libraries 27 Mar 2024 · 1 repository · arXiv:2403.18219
-
DASA: Delay-Adaptive Multi-Agent Stochastic Approximation 25 Mar 2024 · 0 repositories · arXiv:2403.17247
-
Semantic-Aware Remote Estimation of Multiple Markov Sources Under Constraints 25 Mar 2024 · 0 repositories · arXiv:2403.16855
-
A Fairness-Oriented Reinforcement Learning Approach for the Operation and Control of Shared Micromobility Services 23 Mar 2024 · 1 repository · arXiv:2403.15780
-
DouRN: Improving DouZero by Residual Neural Networks 21 Mar 2024 · 0 repositories · arXiv:2403.14102
-
Sparse Bayesian Learning-Based Hierarchical Construction for 3D Radio Environment Maps Incorporating Channel Shadowing 13 Mar 2024 · 0 repositories · arXiv:2403.08323
-
Strategizing against Q-learners: A Control-theoretical Approach 13 Mar 2024 · 0 repositories · arXiv:2403.08906
-
Optimal Design and Implementation of an Open-source Emulation Platform for User-Centric Shared E-mobility Services 12 Mar 2024 · 0 repositories · arXiv:2403.07964
-
Finite-Time Error Analysis of Soft Q-Learning: Switching System Approach 11 Mar 2024 · 0 repositories · arXiv:2403.06366
-
Scalable Online Exploration via Coverability 11 Mar 2024 · 1 repository · arXiv:2403.06571Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Algorithmic Collusion and Price Discrimination: The Over-Usage of Data 10 Mar 2024 · 0 repositories · arXiv:2403.06150
-
Enhancing Classification Performance via Reinforcement Learning for Feature Selection 9 Mar 2024 · 0 repositories · arXiv:2403.05979
-
ComTraQ-MPC: Meta-Trained DQN-MPC Integration for Trajectory Tracking with Limited Active Localization Updates 3 Mar 2024 · 1 repository · arXiv:2403.01564
-
Efficient Episodic Memory Utilization of Cooperative Multi-Agent Reinforcement Learning 2 Mar 2024 · 1 repository · arXiv:2403.01112Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 1 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 4 harvested samples)
-
QF-tuner: Breaking Tradition in Reinforcement Learning 26 Feb 2024 · 0 repositories · arXiv:2402.16562
-
Optimizing Portfolio Management and Risk Assessment in Digital Assets Using Deep Learning for Predictive Analysis 25 Feb 2024 · 0 repositories · arXiv:2402.15994
-
Computation Offloading for Multi-server Multi-access Edge Vehicular Networks: A DDQN-based Method 21 Feb 2024 · 0 repositories · arXiv:2404.07215
-
An Index Policy Based on Sarsa and Q-learning for Heterogeneous Smart Target Tracking 19 Feb 2024 · 0 repositories · arXiv:2402.12015
-
Easy as ABCs: Unifying Boltzmann Q-Learning and Counterfactual Regret Minimization 19 Feb 2024 · 0 repositories · arXiv:2402.11835
-
Interference Mitigation in LEO Constellations with Limited Radio Environment Information 19 Feb 2024 · 0 repositories · arXiv:2402.12103
-
Reinforcement learning to maximise wind turbine energy generation 17 Feb 2024 · 0 repositories · arXiv:2402.11384
-
Enhancing Courier Scheduling in Crowdsourced Last-Mile Delivery through Dynamic Shift Extensions: A Deep Reinforcement Learning Approach 15 Feb 2024 · 0 repositories · arXiv:2402.09961
-
Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks 14 Feb 2024 · 0 repositories · arXiv:2402.09078
-
Conservative and Risk-Aware Offline Multi-Agent Reinforcement Learning 13 Feb 2024 · 1 repository · arXiv:2402.08421Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples)
-
Enhanced Deep Q-Learning for 2D Self-Driving Cars: Implementation and Evaluation on a Custom Track Environment 13 Feb 2024 · 0 repositories · arXiv:2402.08780
-
Intelligent Agricultural Management Considering N₂O Emission and Climate Variability with Uncertainties 13 Feb 2024 · 0 repositories · arXiv:2402.08832
-
Leveraging Digital Cousins for Ensemble Q-Learning in Large-Scale Wireless Networks 12 Feb 2024 · 1 repository · arXiv:2402.08022
-
ORIENT: A Priority-Aware Energy-Efficient Approach for Latency-Sensitive Applications in 6G 10 Feb 2024 · 0 repositories · arXiv:2402.06931
-
Solving Deep Reinforcement Learning Tasks with Evolution Strategies and Linear Policy Networks 10 Feb 2024 · 1 repository · arXiv:2402.06912
-
RLEEGNet: Integrating Brain-Computer Interfaces with Adaptive AI for Intuitive Responsiveness and High-Accuracy Motor Imagery Classification 9 Feb 2024 · 0 repositories · arXiv:2402.09465
-
Value function interference and greedy action selection in value-based multi-objective reinforcement learning 9 Feb 2024 · 0 repositories · arXiv:2402.06266
-
Enhancement of High-definition Map Update Service Through Coverage-aware and Reinforcement Learning 8 Feb 2024 · 0 repositories · arXiv:2402.14582
-
Federated Offline Reinforcement Learning: Collaborative Single-Policy Coverage Suffices 8 Feb 2024 · 0 repositories · arXiv:2402.05876
-
Improving Token-Based World Models with Parallel Observation Prediction 8 Feb 2024 · 1 repository · arXiv:2402.05643Syntology official (archive's flag): 8 ran · 8 ran (of which 0 constructed an object rather than computing a result; 8 with no instrument failure: 0 honoured, 0 violated, 8 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Multi-Timescale Ensemble Q-learning for Markov Decision Process Policy Optimization 8 Feb 2024 · 1 repository · arXiv:2402.05476
-
A Deep Reinforcement Learning Approach for Adaptive Traffic Routing in Next-gen Networks 7 Feb 2024 · 0 repositories · arXiv:2402.04515
-
Averaging n-step Returns Reduces Variance in Reinforcement Learning 6 Feb 2024 · 0 repositories · arXiv:2402.03903Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
Logical Specifications-guided Dynamic Task Sampling for Reinforcement Learning Agents 6 Feb 2024 · 1 repository · arXiv:2402.03678
-
Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning 5 Feb 2024 · 0 repositories · arXiv:2402.03570
-
Q-Star Meets Scalable Posterior Sampling: Bridging Theory and Practice via HyperAgent 5 Feb 2024 · 3 repositories · arXiv:2402.10228Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 7 harvested samples)
-
DRL-Based Dynamic Channel Access and SCLAR Maximization for Networks Under Jamming 2 Feb 2024 · 0 repositories · arXiv:2402.01574
-
Deep Robot Sketching: An application of Deep Q-Learning Networks for human-like sketching 1 Feb 2024 · 0 repositories · arXiv:2402.00676
-
FM3Q: Factorized Multi-Agent MiniMax Q-Learning for Two-Team Zero-Sum Markov Game 1 Feb 2024 · 0 repositories · arXiv:2402.00738
-
RadDQN: a Deep Q Learning-based Architecture for Finding Time-efficient Minimum Radiation Exposure Pathway 1 Feb 2024 · 1 repository · arXiv:2402.00468
-
Extrinsicaly Rewarded Soft Q Imitation Learning with Discriminator 30 Jan 2024 · 0 repositories · arXiv:2401.16772
-
A comparison of RL-based and PID controllers for 6-DOF swimming robots: hybrid underwater object tracking 29 Jan 2024 · 1 repository · arXiv:2401.16618
-
Emergence of cooperation under punishment: A reinforcement learning perspective 29 Jan 2024 · 0 repositories · arXiv:2401.16073
-
Regularized Q-Learning with Linear Function Approximation 26 Jan 2024 · 0 repositories · arXiv:2401.15196
-
Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation 25 Jan 2024 · 0 repositories · arXiv:2401.13884
-
Tacit algorithmic collusion in deep reinforcement learning guided price competition: A study using EV charge pricing game 25 Jan 2024 · 0 repositories · arXiv:2401.15108
-
VQC-Based Reinforcement Learning with Data Re-uploading: Performance and Trainability 21 Jan 2024 · 1 repository · arXiv:2401.11555
-
Attention-Based CNN-BiLSTM for Sleep State Classification of Spatiotemporal Wide-Field Calcium Imaging Data 16 Jan 2024 · 1 repository · arXiv:2401.08098
-
REValueD: Regularised Ensemble Value-Decomposition for Factorisable Markov Decision Processes 16 Jan 2024 · 0 repositories · arXiv:2401.08850
-
A Semantic-Aware Multiple Access Scheme for Distributed, Dynamic 6G-Based Applications 12 Jan 2024 · 1 repository · arXiv:2401.06308
-
Graph Q-Learning for Combinatorial Optimization 11 Jan 2024 · 0 repositories · arXiv:2401.05610
-
Model-Free Reinforcement Learning for Automated Fluid Administration in Critical Care 11 Jan 2024 · 0 repositories · arXiv:2401.06299
-
Deep Reinforcement Multi-agent Learning framework for Information Gathering with Local Gaussian Processes for Water Monitoring 9 Jan 2024 · 0 repositories · arXiv:2401.04631
-
Inverse-like Antagonistic Scene Text Spotting via Reading-Order Estimation and Dynamic Sampling 8 Jan 2024 · 0 repositories · arXiv:2401.03637
-
An Empirical Investigation of Value-Based Multi-objective Reinforcement Learning for Stochastic Environments 6 Jan 2024 · 0 repositories · arXiv:2401.03163
-
Semi-supervised learning via DQN for log anomaly detection 6 Jan 2024 · 0 repositories · arXiv:2401.03151
-
SPQR: Controlling Q-ensemble Independence with Spiked Random Model for Reinforcement Learning 6 Jan 2024 · 1 repository · arXiv:2401.03137Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples)
-
A Deep Q-Learning based Smart Scheduling of EVs for Demand Response in Smart Grids 5 Jan 2024 · 0 repositories · arXiv:2401.02653
-
The Best Time for an Update: Risk-Sensitive Minimization of Age-Based Metrics 3 Jan 2024 · 0 repositories · arXiv:2401.10265
-
Dynamic Decision Making in Engineering System Design: A Deep Q-Learning Approach 28 Dec 2023 · 0 repositories · arXiv:2312.17284
-
Reinforcement Learning for Safe Occupancy Strategies in Educational Spaces during an Epidemic 23 Dec 2023 · 0 repositories · arXiv:2312.15163
-
Federated Q-Learning: Linear Regret Speedup with Low Communication Cost 22 Dec 2023 · 0 repositories · arXiv:2312.15023
-
Optimal coordination of resources: A solution from reinforcement learning 20 Dec 2023 · 0 repositories · arXiv:2312.14970
-
Investigating the Performance and Reliability, of the Q-Learning Algorithm in Various Unknown Environments 19 Dec 2023 · 1 repository
-
Modeling non-linear Effects with Neural Networks in Relational Event Models 19 Dec 2023 · 1 repository · arXiv:2312.12357
-
Sample Efficient Reinforcement Learning with Partial Dynamics Knowledge 19 Dec 2023 · 1 repository · arXiv:2312.12558
-
Stability of Multi-Agent Learning in Competitive Networks: Delaying the Onset of Chaos 19 Dec 2023 · 0 repositories · arXiv:2312.11943
-
Deep-Dispatch: A Deep Reinforcement Learning-Based Vehicle Dispatch Algorithm for Advanced Air Mobility 17 Dec 2023 · 0 repositories · arXiv:2312.10809
-
On Designing Multi-UAV aided Wireless Powered Dynamic Communication via Hierarchical Deep Reinforcement Learning 13 Dec 2023 · 0 repositories · arXiv:2312.07917
-
The Effective Horizon Explains Deep RL Performance in Stochastic Environments 13 Dec 2023 · 1 repository · arXiv:2312.08369
-
Enhanced Q-Learning Approach to Finite-Time Reachability with Maximum Probability for Probabilistic Boolean Control Networks 12 Dec 2023 · 0 repositories · arXiv:2312.06904
-
I Open at the Close: A Deep Reinforcement Learning Evaluation of Open Streets Initiatives 12 Dec 2023 · 1 repository · arXiv:2312.07680
-
Efficient Sparse-Reward Goal-Conditioned Reinforcement Learning with a High Replay Ratio and Regularization 10 Dec 2023 · 1 repository · arXiv:2312.05787
-
Synthesis of Temporally-Robust Policies for Signal Temporal Logic Tasks using Reinforcement Learning 10 Dec 2023 · 1 repository · arXiv:2312.05764
-
Multi-Agent Reinforcement Learning via Distributed MPC as a Function Approximator 8 Dec 2023 · 1 repository · arXiv:2312.05166
-
Efficient Parallel Reinforcement Learning Framework using the Reactor Model 7 Dec 2023 · 1 repository · arXiv:2312.04704
-
An efficient data-based off-policy Q-learning algorithm for optimal output feedback control of linear systems 6 Dec 2023 · 0 repositories · arXiv:2312.03451
-
Wake-Sleep Consolidated Learning 6 Dec 2023 · 0 repositories · arXiv:2401.08623
-
Lights out: training RL agents robust to temporary blindness 5 Dec 2023 · 0 repositories · arXiv:2312.02665
-
Provable Reinforcement Learning for Networked Control Systems with Stochastic Packet Disordering 5 Dec 2023 · 0 repositories · arXiv:2312.02498
-
Algorithmic collusion under competitive design 5 Dec 2023 · 0 repositories · arXiv:2312.02644
-
AdsorbRL: Deep Multi-Objective Reinforcement Learning for Inverse Catalysts Design 4 Dec 2023 · 1 repository · arXiv:2312.02308Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Anomaly Detection via Learning-Based Sequential Controlled Sensing 30 Nov 2023 · 0 repositories · arXiv:2312.00088
-
Data-efficient Deep Reinforcement Learning for Vehicle Trajectory Control 30 Nov 2023 · 0 repositories · arXiv:2311.18393