Browse State-of-the-Art › Q-Learning › Papers, page 14
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 14 of 20: papers 1,301 to 1,400 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Reinforcement Learning for Traffic Signal Control: Comparison with Commercial Systems21 Apr 2021 0 repositories listed
-
A Simulated Experiment to Explore Robotic Dialogue Strategies for People with Dementia18 Apr 2021 0 repositories listed
-
Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills15 Apr 2021 0 repositories listed
-
Prospect-theoretic Q-learning12 Apr 2021 0 repositories listed
-
Towards Resilience for Multi-Agent QD-Learning7 Apr 2021 0 repositories listed
-
Distributed Deep Reinforcement Learning for Collaborative Spectrum Sharing6 Apr 2021 0 repositories listed
-
SOLO: Search Online, Learn Offline for Combinatorial Optimization Problems4 Apr 2021 0 repositories listed
-
Federated Double Deep Q-learning for Joint Delay and Energy Minimization in IoT networks2 Apr 2021 0 repositories listed
-
Convergence of Finite Memory Q-Learning for POMDPs and Near Optimality of Learned Policies under Filter Stability22 Mar 2021 0 repositories listed
-
Reinforcement Learning based on Scenario-tree MPC for ASVs22 Mar 2021 0 repositories listed
-
Regularized Softmax Deep Multi-Agent Q-Learning22 Mar 2021 0 repositories listed
-
Variational quantum compiling with double Q-learning22 Mar 2021 0 repositories listed
-
A Jointly Optimal Design of Control and Scheduling in Networked Systems under Denial-of-Service Attacks10 Mar 2021 0 repositories listed
-
S4RL: Surprisingly Simple Self-Supervision for Offline Reinforcement Learning10 Mar 2021 0 repositories listed
-
The Effect of Q-function Reuse on the Total Regret of Tabular, Model-Free, Reinforcement Learning7 Mar 2021 0 repositories listed
-
Correlated Deep Q-learning based Microgrid Energy Management6 Mar 2021 0 repositories listed
-
Decentralized Microgrid Energy Management: A Multi-agent Correlated Q-learning Approach6 Mar 2021 0 repositories listed
-
Ensemble Bootstrapping for Q-Learning28 Feb 2021 0 repositories listed
-
Potential Impacts of Smart Homes on Human Behavior: A Reinforcement Learning Approach26 Feb 2021 0 repositories listed
-
No-Regret Reinforcement Learning with Heavy-Tailed Rewards25 Feb 2021 0 repositories listed
-
Reinforcement learning approach for resource allocation in humanitarian logistics25 Feb 2021 0 repositories listed
-
Sequential Learning-based IaaS Composition24 Feb 2021 0 repositories listed
-
Greedy-Step Off-Policy Reinforcement Learning23 Feb 2021 0 repositories listed
-
A Discrete-Time Switching System Analysis of Q-learning17 Feb 2021 0 repositories listed
-
Cooperation and Reputation Dynamics with Reinforcement Learning15 Feb 2021 0 repositories listed
-
Reversible Action Design for Combinatorial Optimization with Reinforcement Learning14 Feb 2021 0 repositories listed
-
Is Q-Learning Minimax Optimal? A Tight Sample Complexity Analysis12 Feb 2021 0 repositories listed
-
Hedging of Financial Derivative Contracts via Monte Carlo Tree Search11 Feb 2021 0 repositories listed
-
Simple Agent, Complex Environment: Efficient Reinforcement Learning with Agent States10 Feb 2021 0 repositories listed
-
Model-Augmented Q-learning7 Feb 2021 0 repositories listed
-
Experience-Based Heuristic Search: Robust Motion Planning with Deep Q-Learning5 Feb 2021 0 repositories listed
-
A review of motion planning algorithms for intelligent robotics4 Feb 2021 0 repositories listed
-
Deep reinforcement learning-based image classification achieves perfect testing set accuracy for MRI brain tumors with a training set of only 30 images4 Feb 2021 0 repositories listed
-
A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants2 Feb 2021 0 repositories listed
-
QoS-Aware Power Minimization of Distributed Many-Core Servers using Transfer Q-Learning2 Feb 2021 0 repositories listed
-
A step toward a reinforcement learning de novo genome assembler2 Feb 2021 0 repositories listed
-
CoordiQ : Coordinated Q-learning for Electric Vehicle Charging Recommendation28 Jan 2021 0 repositories listed
-
Reinforcement Learning based Per-antenna Discrete Power Control for Massive MIMO Systems28 Jan 2021 0 repositories listed
-
Reinforcement Learning Assisted Beamforming for Inter-cell Interference Mitigation in 5G Massive MIMO Networks27 Jan 2021 0 repositories listed
-
Robust Android Malware Detection System against Adversarial Attacks using Q-Learning27 Jan 2021 0 repositories listed
-
Channel Estimation via Successive Denoising in MIMO OFDM Systems: A Reinforcement Learning Approach25 Jan 2021 0 repositories listed
-
Solving optimal stopping problems with Deep Q-Learning24 Jan 2021 0 repositories listed
-
Fire Threat Detection From Videos with Q-Rough Sets21 Jan 2021 0 repositories listed
-
Reinforcement learning based recommender systems: A survey15 Jan 2021 0 repositories listed
-
Learning Augmented Index Policy for Optimal Service Placement at the Network Edge10 Jan 2021 0 repositories listed
-
Robust and Scalable Routing with Multi-Agent Deep Reinforcement Learning for MANETs9 Jan 2021 0 repositories listed
-
Safe Coupled Deep Q-Learning for Recommendation Systems8 Jan 2021 0 repositories listed
-
Addressing Distribution Shift in Online Reinforcement Learning with Offline Datasets1 Jan 2021 0 repositories listed
-
Deep Q Learning from Dynamic Demonstration with Behavioral Cloning1 Jan 2021 0 repositories listed
-
Deep Q-Learning with Low Switching Cost1 Jan 2021 0 repositories listed
-
Deep Reinforcement Learning-based Anti-jamming Power Allocation in a Two-cell NOMA Network1 Jan 2021 0 repositories listed
-
Double Q-learning: New Analysis and Sharper Finite-time Bound1 Jan 2021 0 repositories listed
-
Learning Movement Strategies for Moving Target Defense1 Jan 2021 0 repositories listed
-
Optimistic Exploration with Backward Bootstrapped Bonus for Deep Reinforcement Learning1 Jan 2021 0 repositories listed
-
Preventing Value Function Collapse in Ensemble Q-Learning by Maximizing Representation Diversity1 Jan 2021 0 repositories listed
-
Success-Rate Targeted Reinforcement Learning by Disorientation Penalty1 Jan 2021 0 repositories listed
-
Uncertainty Weighted Offline Reinforcement Learning1 Jan 2021 0 repositories listed
-
Weighted Bellman Backups for Improved Signal-to-Noise in Q-Updates1 Jan 2021 0 repositories listed
-
Blackwell Online Learning for Markov Decision Processes28 Dec 2020 0 repositories listed
-
Disentangled Planning and Control in Vision Based Robotics via Reward Machines28 Dec 2020 0 repositories listed
-
Assured RL: Reinforcement Learning with Almost Sure Constraints24 Dec 2020 0 repositories listed
-
Distributed Q-Learning with State Tracking for Multi-agent Networked Control22 Dec 2020 0 repositories listed
-
Goal Reasoning by Selecting Subgoals with Deep Q-Learning22 Dec 2020 0 repositories listed
-
Stabilizing Q Learning Via Soft Mellowmax Operator17 Dec 2020 0 repositories listed
-
Sample-Efficient Reinforcement Learning via Counterfactual-Based Data Augmentation16 Dec 2020 0 repositories listed
-
Deploying Reinforcement Learning in Water Transport14 Dec 2020 0 repositories listed
-
Virtual Autonomous Driving with Reinforcement Learning14 Dec 2020 0 repositories listed
-
Semi-Supervised Off Policy Reinforcement Learning9 Dec 2020 0 repositories listed
-
Selective Pseudo-Labeling with Reinforcement Learning for Semi-Supervised Domain Adaptation7 Dec 2020 0 repositories listed
-
Amortized Q-learning with Model-based Action Proposals for Autonomous Driving on Highways6 Dec 2020 0 repositories listed
-
Hippocampal representations emerge when training recurrent neural networks on a memory dependent maze navigation task2 Dec 2020 0 repositories listed
-
Self-correcting Q-Learning2 Dec 2020 0 repositories listed
-
A new convergent variant of Q-learning with linear function approximation1 Dec 2020 0 repositories listed
-
A Unified Switching System Perspective and Convergence Analysis of Q-Learning Algorithms1 Dec 2020 0 repositories listed
-
Agnostic Q-learning with Function Approximation in Deterministic Systems: Near-Optimal Bounds on Approximation Error and Sample Complexity1 Dec 2020 0 repositories listed
-
Can Temporal-Difference and Q-Learning Learn Representation? A Mean-Field Theory1 Dec 2020 0 repositories listed
-
Robust Multi-Agent Reinforcement Learning with Model Uncertainty1 Dec 2020 0 repositories listed
-
Deep reinforcement learning with a particle dynamics environment applied to emergency evacuation of a room with obstacles30 Nov 2020 0 repositories listed
-
Real-time Active Vision for a Humanoid Soccer Robot Using Deep Reinforcement Learning27 Nov 2020 0 repositories listed
-
Reinforcement Learning-based Joint Path and Energy Optimization of Cellular-Connected Unmanned Aerial Vehicles27 Nov 2020 0 repositories listed
-
Diluted Near-Optimal Expert Demonstrations for Guiding Dialogue Stochastic Policy Optimisation25 Nov 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning for Markov Routing Games: A New Modeling Paradigm For Dynamic Traffic Assignment22 Nov 2020 0 repositories listed
-
Provable Multi-Objective Reinforcement Learning with Generative Models19 Nov 2020 0 repositories listed
-
C-Learning: Learning to Achieve Goals via Recursive Classification17 Nov 2020 0 repositories listed
-
Constrained Model-Free Reinforcement Learning for Process Optimization16 Nov 2020 0 repositories listed
-
A deep Q-Learning based Path Planning and Navigation System for Firefighting Environments12 Nov 2020 0 repositories listed
-
On Using Hamiltonian Monte Carlo Sampling for Reinforcement Learning Problems in High-dimension11 Nov 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning for Channel Assignment and Power Allocation in Platoon-Based C-V2X Systems9 Nov 2020 0 repositories listed
-
Reinforced Deep Markov Models With Applications in Automatic Trading9 Nov 2020 0 repositories listed
-
Reinforcement Learning for Assignment problem8 Nov 2020 0 repositories listed
-
A Hysteretic Q-learning Coordination Framework for Emerging Mobility Systems in Smart Cities5 Nov 2020 0 repositories listed
-
DeepFoldit -- A Deep Reinforcement Learning Neural Network Folding Proteins28 Oct 2020 0 repositories listed
-
Finite-Time Convergence Rates of Decentralized Stochastic Approximation with Applications in Multi-Agent and Multi-Task Learning28 Oct 2020 0 repositories listed
-
Energy Consumption and Battery Aging Minimization Using a Q-learning Strategy for a Battery/Ultracapacitor Electric Vehicle27 Oct 2020 0 repositories listed
-
Learning Time Reduction Using Warm Start Methods for a Reinforcement Learning Based Supervisory Control in Hybrid Electric Vehicle Applications27 Oct 2020 0 repositories listed
-
Energy and Service-priority aware Trajectory Design for UAV-BSs using Double Q-Learning26 Oct 2020 0 repositories listed
-
Enhancing reinforcement learning by a finite reward response filter with a case study in intelligent structural control25 Oct 2020 0 repositories listed
-
An Adiabatic Theorem for Policy Tracking with TD-learning24 Oct 2020 0 repositories listed
-
Stabilizing Transformer-Based Action Sequence Generation For Q-Learning23 Oct 2020 0 repositories listed
-
Deep Surrogate Q-Learning for Autonomous Driving21 Oct 2020 0 repositories listed