Methods › Reinforcement Learning › Off-Policy TD Control › Q-Learning › Papers, page 5
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,734 · with a code link: 464 · where Syntology ran a sample: 126 (105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (126 of 1,734 tagged: 105 with a run with no instrument failure, 21 where every run was a failure of Syntology's instrument)
Page 5 of 18: papers 401 to 500 of 1,734, newest first by the archive's date (ties by slug), in archive order.
Repository counts are the archive's code-links table. A Syntology line states what Syntology ran from that paper's harvested code, as “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the instrument figure counts failures of Syntology's instrument, not of the code. It is per sample and not a correctness claim. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code; “community repositories only” when every sample that ran came from a community repository, “official: no sample here; runs from other or unrecorded repositories” when some came from a repository the paper names or has in its text, or from none recorded); hover it for the repositories the samples that ran came from.
-
Q-learning Based Optimal False Data Injection Attack on Probabilistic Boolean Control Networks 29 Nov 2023 · 0 repositories · arXiv:2311.17631
-
Self-Driving Telescopes: Autonomous Scheduling of Astronomical Observation Campaigns with Offline Reinforcement Learning 29 Nov 2023 · 0 repositories · arXiv:2311.18094
-
Reinforcement Learning for Wildfire Mitigation in Simulated Disaster Environments 27 Nov 2023 · 1 repository · arXiv:2311.15925Syntology official (archive's flag): 5 ran · 5 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 7 harvested samples)
-
Reinforcement Learning from Diffusion Feedback: Q* for Image Search 27 Nov 2023 · 0 repositories · arXiv:2311.15648
-
A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation 26 Nov 2023 · 0 repositories · arXiv:2311.15238
-
Projected Off-Policy Q-Learning (POP-QL) for Stabilizing Offline Reinforcement Learning 25 Nov 2023 · 0 repositories · arXiv:2311.14885
-
Approximation of Convex Envelope Using Reinforcement Learning 24 Nov 2023 · 0 repositories · arXiv:2311.14421
-
Efficient Open-world Reinforcement Learning via Knowledge Distillation and Autonomous Rule Discovery 24 Nov 2023 · 0 repositories · arXiv:2311.14270
-
On optimal tracking portfolio in incomplete markets: The reinforcement learning approach 24 Nov 2023 · 0 repositories · arXiv:2311.14318
-
Multi-intention Inverse Q-learning for Interpretable Behavior Representation 23 Nov 2023 · 1 repository · arXiv:2311.13870
-
Machine learning-based decentralized TDMA for VLC IoT networks 23 Nov 2023 · 0 repositories · arXiv:2311.14078
-
Decentralised Q-Learning for Multi-Agent Markov Decision Processes with a Satisfiability Criterion 21 Nov 2023 · 0 repositories · arXiv:2311.12613
-
Offline Reinforcement Learning for Wireless Network Optimization with Mixture Datasets 19 Nov 2023 · 0 repositories · arXiv:2311.11423
-
Genetic Algorithm enhanced by Deep Reinforcement Learning in parent selection mechanism and mutation : Minimizing makespan in permutation flow shop scheduling problems 10 Nov 2023 · 0 repositories · arXiv:2311.05937
-
Two-compartment neuronal spiking model expressing brain-state specific apical-amplification, -isolation and -drive regimes 10 Nov 2023 · 0 repositories · arXiv:2311.06074
-
Advancing Algorithmic Trading: A Multi-Technique Enhancement of Deep Q-Network Models 9 Nov 2023 · 0 repositories · arXiv:2311.05743
-
Selectively Sharing Experiences Improves Multi-Agent Reinforcement Learning 1 Nov 2023 · 1 repository · arXiv:2311.00865Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments 31 Oct 2023 · 0 repositories · arXiv:2311.00123
-
DGFN: Double Generative Flow Networks 30 Oct 2023 · 0 repositories · arXiv:2310.19685
-
Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning 29 Oct 2023 · 0 repositories · arXiv:2310.19137
-
Weakly Coupled Deep Q-Networks 28 Oct 2023 · 0 repositories · arXiv:2310.18803Syntology 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Lifting the Veil: Unlocking the Power of Depth in Q-learning 27 Oct 2023 · 0 repositories · arXiv:2310.17915
-
Model-free Posterior Sampling via Learning Rate Randomization 27 Oct 2023 · 0 repositories · arXiv:2310.18186
-
Integrated Freeway Traffic Control Using Q-Learning with Adjacent Arterial Traffic Considerations 25 Oct 2023 · 0 repositories · arXiv:2310.16748
-
On the Convergence and Sample Complexity Analysis of Deep Q-Networks with ε-Greedy Exploration 24 Oct 2023 · 0 repositories · arXiv:2310.16173
-
Reinforcement learning based local path planning for mobile robot 24 Oct 2023 · 0 repositories · arXiv:2403.12463
-
AI on the Water: Applying DRL to Autonomous Vessel Navigation 23 Oct 2023 · 0 repositories · arXiv:2310.14938
-
Deep Reinforcement Learning-based Intelligent Traffic Signal Controls with Optimized CO2 emissions 19 Oct 2023 · 1 repository · arXiv:2310.13129
-
Towards Robust Offline Reinforcement Learning under Diverse Data Corruption 19 Oct 2023 · 2 repositories · arXiv:2310.12955Syntology official (archive's flag): 2 ran · 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Automatic Music Playlist Generation via Simulation-based Reinforcement Learning 13 Oct 2023 · 0 repositories · arXiv:2310.09123
-
Optimal Scheduling of Electric Vehicle Charging with Deep Reinforcement Learning considering End Users Flexibility 13 Oct 2023 · 0 repositories · arXiv:2310.09040
-
Bad Values but Good Behavior: Learning Highly Misspecified Bandits and MDPs 13 Oct 2023 · 0 repositories · arXiv:2310.09358
-
Learning RL-Policies for Joint Beamforming Without Exploration: A Batch Constrained Off-Policy Approach 12 Oct 2023 · 1 repository · arXiv:2310.08660
-
Integrated Sensing and Communication Neighbor Discovery for MANET with Gossip Mechanism 11 Oct 2023 · 0 repositories · arXiv:2310.07292
-
Boosting Continuous Control with Consistency Policy 10 Oct 2023 · 1 repository · arXiv:2310.06343Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Suppressing Overestimation in Q-Learning through Adversarial Behaviors 10 Oct 2023 · 0 repositories · arXiv:2310.06286
-
DeepQTest: Testing Autonomous Driving Systems with Reinforcement Learning and Real-world Weather Data 8 Oct 2023 · 1 repository · arXiv:2310.05170
-
Digital Twin Assisted Deep Reinforcement Learning for Online Admission Control in Sliced Network 7 Oct 2023 · 0 repositories · arXiv:2310.09299
-
Optimal Sequential Decision-Making in Geosteering: A Reinforcement Learning Approach 7 Oct 2023 · 0 repositories · arXiv:2310.04772
-
Applying Reinforcement Learning to Option Pricing and Hedging 6 Oct 2023 · 0 repositories · arXiv:2310.04336
-
Optimal Control of District Cooling Energy Plant with Reinforcement Learning and MPC 5 Oct 2023 · 0 repositories · arXiv:2310.03814
-
A Deep Reinforcement Learning Approach for Interactive Search with Sentence-level Feedback 3 Oct 2023 · 0 repositories · arXiv:2310.03043
-
Finite-Time Analysis of Whittle Index based Q-Learning for Restless Multi-Armed Bandits with Neural Network Function Approximation 3 Oct 2023 · 0 repositories · arXiv:2310.02147
-
PGDQN: Preference-Guided Deep Q-Network 3 Oct 2023 · 1 repository
-
Using Reinforcement Learning to Optimize Responses in Care Processes: A Case Study on Aggression Incidents 2 Oct 2023 · 0 repositories · arXiv:2310.00981
-
Pre-training with Synthetic Data Helps Offline Reinforcement Learning 1 Oct 2023 · 1 repository · arXiv:2310.00771Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 7 harvested samples)
-
Reinforcement learning adaptive fuzzy controller for lighting systems: application to aircraft cabin 30 Sep 2023 · 0 repositories · arXiv:2310.00525
-
SI-SD: Sleep Interpreter through awake-guided cross-subject Semantic Decoding 28 Sep 2023 · 0 repositories · arXiv:2309.16457
-
Decoding trust: A reinforcement learning perspective 26 Sep 2023 · 0 repositories · arXiv:2309.14598
-
Adapting Double Q-Learning for Continuous Reinforcement Learning 25 Sep 2023 · 0 repositories · arXiv:2309.14471
-
Deep Reinforcement Learning for the Heat Transfer Control of Pulsating Impinging Jets 25 Sep 2023 · 0 repositories · arXiv:2309.13955
-
Enhancing data efficiency in reinforcement learning: a novel imagination mechanism based on mesh information propagation 25 Sep 2023 · 2 repositories · arXiv:2309.14243
-
Enhancing Healthcare with EOG: A Novel Approach to Sleep Stage Classification 25 Sep 2023 · 0 repositories · arXiv:2310.03757
-
Implicit Sensing in Traffic Optimization: Advanced Deep Reinforcement Learning Techniques 25 Sep 2023 · 0 repositories · arXiv:2309.14395
-
Counterfactual Conservative Q Learning for Offline Multi-agent Reinforcement Learning 22 Sep 2023 · 1 repository · arXiv:2309.12696Syntology official (archive's flag): 6 ran · 7 ran (of which 3 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 3 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Belief Projection-Based Reinforcement Learning for Environments with Delayed Feedback 21 Sep 2023 · 1 repository
-
Double Gumbel Q-Learning 21 Sep 2023 · 1 repository
-
On the Convergence and Sample Complexity Analysis of Deep Q-Networks with ϵ-Greedy Exploration 21 Sep 2023 · 0 repositories
-
ReDS: Offline RL With Heteroskedastic Datasets via Support Constraints 21 Sep 2023 · 0 repositories
-
UAV Swarm Deployment and Trajectory for 3D Area Coverage via Reinforcement Learning 21 Sep 2023 · 0 repositories · arXiv:2309.11992
-
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions 20 Sep 2023 · 0 repositories · arXiv:2309.10980
-
Differentiable Quantum Architecture Search for Quantum Reinforcement Learning 19 Sep 2023 · 0 repositories · arXiv:2309.10392
-
Double Deep Q-Learning-based Path Selection and Service Placement for Latency-Sensitive Beyond 5G Applications 18 Sep 2023 · 0 repositories · arXiv:2309.10180
-
Self-Sustaining Multiple Access with Continual Deep Reinforcement Learning for Dynamic Metaverse Applications 18 Sep 2023 · 0 repositories · arXiv:2309.10177
-
Dynamic control of self-assembly of quasicrystalline structures through reinforcement learning 13 Sep 2023 · 2 repositories · arXiv:2309.06869
-
A Q-learning Approach for Adherence-Aware Recommendations 12 Sep 2023 · 0 repositories · arXiv:2309.06519
-
Career Path Recommendations for Long-term Income Maximization: A Reinforcement Learning Approach 11 Sep 2023 · 0 repositories · arXiv:2309.05391
-
Convex Q Learning in a Stochastic Environment: Extended Version 10 Sep 2023 · 0 repositories · arXiv:2309.05105
-
Multi Agent DeepRL based Joint Power and Subchannel Allocation in IAB networks 31 Aug 2023 · 0 repositories · arXiv:2309.00144
-
Coalescent processes emerging from large deviations 28 Aug 2023 · 0 repositories · arXiv:2308.14715
-
Learning Visual Tracking and Reaching with Deep Reinforcement Learning on a UR10e Robotic Arm 28 Aug 2023 · 1 repository · arXiv:2308.14652
-
Prompt to Transfer: Sim-to-Real Transfer for Traffic Signal Control with Prompt Learning 28 Aug 2023 · 1 repository · arXiv:2308.14284Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
Reinforcement Learning for Sampling on Temporal Medical Imaging Sequences 28 Aug 2023 · 1 repository · arXiv:2308.14946Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
A Graph Neural Network-Based QUBO-Formulated Hamiltonian-Inspired Loss Function for Combinatorial Optimization using Reinforcement Learning 27 Aug 2023 · 0 repositories · arXiv:2308.13978
-
Actuator Trajectory Planning for UAVs with Overhead Manipulator using Reinforcement Learning 24 Aug 2023 · 0 repositories · arXiv:2308.12843
-
Towards Few-shot Coordination: Revisiting Ad-hoc Teamplay Challenge In the Game of Hanabi 20 Aug 2023 · 1 repository · arXiv:2308.10284Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
Improving Sample Efficiency of Model-Free Algorithms for Zero-Sum Markov Games 17 Aug 2023 · 0 repositories · arXiv:2308.08858
-
Reinforcement Learning for Battery Management in Dairy Farming 17 Aug 2023 · 0 repositories · arXiv:2308.09023
-
Integrating Renewable Energy in Agriculture: A Deep Reinforcement Learning-based Approach 16 Aug 2023 · 0 repositories · arXiv:2308.08611
-
On-demand Cold Start Frequency Reduction with Off-Policy Reinforcement Learning in Serverless Computing 15 Aug 2023 · 0 repositories · arXiv:2308.07541
-
Variations on the Reinforcement Learning performance of Blackjack 9 Aug 2023 · 2 repositories · arXiv:2308.07329
-
Unsynchronized Decentralized Q-Learning: Two Timescale Analysis By Persistence 7 Aug 2023 · 0 repositories · arXiv:2308.03239
-
Deep Q-Network for Stochastic Process Environments 7 Aug 2023 · 0 repositories · arXiv:2308.03316
-
Bag of Policies for Distributional Deep Exploration 3 Aug 2023 · 0 repositories · arXiv:2308.01759
-
Caching-at-STARS: the Next Generation Edge Caching 1 Aug 2023 · 0 repositories · arXiv:2308.00562
-
Pixel to policy: DQN Encoders for within & cross-game reinforcement learning 1 Aug 2023 · 0 repositories · arXiv:2308.00318
-
Robust Multi-Agent Reinforcement Learning with State Uncertainty 30 Jul 2023 · 1 repository · arXiv:2307.16212Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 1 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 3 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
ETHER: Aligning Emergent Communication for Hindsight Experience Replay 28 Jul 2023 · 0 repositories · arXiv:2307.15494
-
Stability of Multi-Agent Learning: Convergence in Network Games with Many Players 26 Jul 2023 · 0 repositories · arXiv:2307.13922
-
A Flexible Framework for Incorporating Patient Preferences Into Q-Learning 22 Jul 2023 · 0 repositories · arXiv:2307.12022
-
Exploring reinforcement learning techniques for discrete and continuous control tasks in the MuJoCo environment 20 Jul 2023 · 1 repository · arXiv:2307.11166
-
Goal-Conditioned Reinforcement Learning with Disentanglement-based Reachability Planning 20 Jul 2023 · 0 repositories · arXiv:2307.10846
-
Distributed 3D-Beam Reforming for Hovering-Tolerant UAVs Communication over Coexistence: A Deep-Q Learning for Intelligent Space-Air-Ground Integrated Networks 18 Jul 2023 · 0 repositories · arXiv:2307.09325
-
Meta-Value Learning: a General Framework for Learning with Learning Awareness 17 Jul 2023 · 1 repository · arXiv:2307.08863
-
Credit Assignment: Challenges and Opportunities in Developing Human-like AI Agents 16 Jul 2023 · 0 repositories · arXiv:2307.08171
-
Deep reinforcement learning for the dynamic vehicle dispatching problem: An event-based approach 13 Jul 2023 · 0 repositories · arXiv:2307.07508
-
Realtime Spectrum Monitoring via Reinforcement Learning -- A Comparison Between Q-Learning and Heuristic Methods 11 Jul 2023 · 0 repositories · arXiv:2307.05763
-
Measuring and Mitigating Interference in Reinforcement Learning 10 Jul 2023 · 0 repositories · arXiv:2307.04887
-
Probabilistic Counterexample Guidance for Safer Reinforcement Learning (Extended Version) 10 Jul 2023 · 1 repository · arXiv:2307.04927
-
Investigating the Edge of Stability Phenomenon in Reinforcement Learning 9 Jul 2023 · 0 repositories · arXiv:2307.04210