Browse State-of-the-Art › Q-Learning › Papers, page 15
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 15 of 20: papers 1,401 to 1,500 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Logistic Q-Learning21 Oct 2020 0 repositories listed
-
On Information Asymmetry in Competitive Multi-Agent Reinforcement Learning: Convergence and Optimality21 Oct 2020 0 repositories listed
-
Reinforcement learning using Deep Q Networks and Q learning accurately localizes brain tumors on MRI with very small training sets21 Oct 2020 0 repositories listed
-
Language Inference with Multi-head Automata through Reinforcement Learning20 Oct 2020 0 repositories listed
-
Learning Dexterous Manipulation from Suboptimal Experts16 Oct 2020 0 repositories listed
-
A Nesterov's Accelerated quasi-Newton method for Global Routing using Deep Reinforcement Learning15 Oct 2020 0 repositories listed
-
Model-Based Reinforcement Learning for Type 1Diabetes Blood Glucose Control13 Oct 2020 0 repositories listed
-
Parameterized Reinforcement Learning for Optical System Optimization9 Oct 2020 0 repositories listed
-
Fictitious play in zero-sum stochastic games8 Oct 2020 0 repositories listed
-
Model-Free Non-Stationary RL: Near-Optimal Regret and Applications in Multi-Agent RL and Inventory Control7 Oct 2020 0 repositories listed
-
Machine Learning Empowered Trajectory and Passive Beamforming Design in UAV-RIS Wireless Networks6 Oct 2020 0 repositories listed
-
Cross Learning in Deep Q-Networks29 Sep 2020 0 repositories listed
-
Finite-Time Analysis for Double Q-learning29 Sep 2020 0 repositories listed
-
Deep Jump Q-Evaluation for Offline Policy Evaluation in Continuous Action Space28 Sep 2020 0 repositories listed
-
Near-Optimal Regret Bounds for Model-Free RL in Non-Stationary Episodic MDPs28 Sep 2020 0 repositories listed
-
Towards Understanding Linear Value Decomposition in Cooperative Multi-Agent Q-Learning28 Sep 2020 0 repositories listed
-
A New Approach for Tactical Decision Making in Lane Changing: Sample Efficient Deep Q Learning with a Safety Feedback Reward24 Sep 2020 0 repositories listed
-
Is Q-Learning Provably Efficient? An Extended Analysis22 Sep 2020 0 repositories listed
-
Hidden Incentives for Auto-Induced Distributional Shift19 Sep 2020 0 repositories listed
-
Reinforcement Learning for Dynamic Resource Optimization in 5G Radio Access Network Slicing14 Sep 2020 0 repositories listed
-
AoI Minimization in Status Update Control with Energy Harvesting Sensors9 Sep 2020 0 repositories listed
-
Deep Reinforcement Learning for Option Replication and Hedging9 Sep 2020 0 repositories listed
-
A Hybrid PAC Reinforcement Learning Algorithm5 Sep 2020 0 repositories listed
-
PAC Reinforcement Learning Algorithm for General-Sum Markov Games5 Sep 2020 0 repositories listed
-
Using Machine Teaching to Investigate Human Assumptions when Teaching Reinforcement Learners5 Sep 2020 0 repositories listed
-
Learning Nash Equilibria in Zero-Sum Stochastic Games via Entropy-Regularized Policy Approximation1 Sep 2020 0 repositories listed
-
Solving the single-track train scheduling problem via Deep Reinforcement Learning1 Sep 2020 0 repositories listed
-
Inverse Policy Evaluation for Value-based Sequential Decision-making26 Aug 2020 0 repositories listed
-
Deep Q-Learning: Theoretical Insights from an Asymptotic Analysis25 Aug 2020 0 repositories listed
-
The reinforcement learning-based multi-agent cooperative approach for the adaptive speed regulation on a metallurgical pickling line16 Aug 2020 0 repositories listed
-
Chrome Dino Run using Reinforcement Learning15 Aug 2020 0 repositories listed
-
Decision-making at Unsignalized Intersection for Autonomous Vehicles: Left-turn Maneuver with Deep Reinforcement Learning14 Aug 2020 0 repositories listed
-
Multi-Agent Double Deep Q-Learning for Beamforming in mmWave MIMO Networks13 Aug 2020 0 repositories listed
-
Caching Placement and Resource Allocation for Cache-Enabling UAV NOMA Networks12 Aug 2020 0 repositories listed
-
Convex Q-Learning, Part 1: Deterministic Optimal Control8 Aug 2020 0 repositories listed
-
Evaluating Load Models and Their Impacts on Power Transfer Limits7 Aug 2020 0 repositories listed
-
Deep Q-Network Based Multi-agent Reinforcement Learning with Binary Action Agents6 Aug 2020 0 repositories listed
-
A Comparative Analysis of Deep Reinforcement Learning-enabled Freeway Decision-making for Automated Vehicles4 Aug 2020 0 repositories listed
-
GenCos' Behaviors Modeling Based on Q Learning Improved by Dichotomy4 Aug 2020 0 repositories listed
-
Cooperative Control of Mobile Robots with Stackelberg Learning3 Aug 2020 0 repositories listed
-
Momentum Q-learning with Finite-Sample Convergence Guarantee30 Jul 2020 0 repositories listed
-
Deep Reinforcement Learning for Dynamic Spectrum Sensing and Aggregation in Multi-Channel Wireless Networks28 Jul 2020 0 repositories listed
-
Variance Reduction for Deep Q-Learning using Stochastic Recursive Gradient25 Jul 2020 0 repositories listed
-
A Comparative Study of AI-based Intrusion Detection Techniques in Critical Infrastructures24 Jul 2020 0 repositories listed
-
Trade-off on Sim2Real Learning: Real-world Learning Faster than Simulations21 Jul 2020 0 repositories listed
-
EMaQ: Expected-Max Q-Learning Operator for Simple Yet Effective Offline and Online RL21 Jul 2020 0 repositories listed
-
A Machine Learning Approach for Task and Resource Allocation in Mobile Edge Computing Based Networks20 Jul 2020 0 repositories listed
-
Multi-agent Reinforcement Learning in Bayesian Stackelberg Markov Games for Adaptive Moving Target Defense20 Jul 2020 0 repositories listed
-
Same-Day Delivery with Fairness19 Jul 2020 0 repositories listed
-
DRIFT: Deep Reinforcement Learning for Functional Software Testing16 Jul 2020 0 repositories listed
-
Meta-Gradient Reinforcement Learning with an Objective Discovered Online16 Jul 2020 0 repositories listed
-
Reinforcement Learning-Enabled Decision-Making Strategies for a Vehicle-Cyber-Physical-System in Connected Environment16 Jul 2020 0 repositories listed
-
Analysis of Q-learning with Adaptation and Momentum Restart for Gradient Descent15 Jul 2020 0 repositories listed
-
Qgraph-bounded Q-learning: Stabilizing Model-Free Off-Policy Deep Reinforcement Learning15 Jul 2020 0 repositories listed
-
Hedging using reinforcement learning: Contextual k-Armed Bandit versus Q-learning3 Jul 2020 0 repositories listed
-
Regularly Updated Deterministic Policy Gradient Algorithm1 Jul 2020 0 repositories listed
-
Provably More Efficient Q-Learning in the One-Sided-Feedback/Full-Feedback Settings30 Jun 2020 0 repositories listed
-
Concept and the implementation of a tool to convert industry 4.0 environments modeled as FSM to an OpenAI Gym wrapper29 Jun 2020 0 repositories listed
-
Using Reinforcement Learning to Herd a Robotic Swarm to a Target Distribution29 Jun 2020 0 repositories listed
-
Active Finite Reward Automaton Inference and Reinforcement Learning Using Queries and Counterexamples28 Jun 2020 0 repositories listed
-
Reinforcement Learning Based Handwritten Digit Recognition with Two-State Q-Learning28 Jun 2020 0 repositories listed
-
Q-Learning with Differential Entropy of Q-Tables26 Jun 2020 0 repositories listed
-
Deep Q-Network-Driven Catheter Segmentation in 3D US by Hybrid Constrained Semi-Supervised Learning and Dual-UNet25 Jun 2020 0 repositories listed
-
Energy Minimization in UAV-Aided Networks: Actor-Critic Learning for Constrained Scheduling Optimization24 Jun 2020 0 repositories listed
-
Preventing Value Function Collapse in Ensemble Q-Learning by Maximizing Representation Diversity24 Jun 2020 0 repositories listed
-
Unified Reinforcement Q-Learning for Mean Field Game and Control Problems24 Jun 2020 0 repositories listed
-
Deep Reinforcement Learning Control for Radar Detection and Tracking in Congested Spectral Environments23 Jun 2020 0 repositories listed
-
Near-Optimal Reinforcement Learning with Self-Play22 Jun 2020 0 repositories listed
-
Risk-Sensitive Reinforcement Learning: Near-Optimal Risk-Sample Tradeoff in Regret22 Jun 2020 0 repositories listed
-
Hybridizing the 1/5-th Success Rule with Q-Learning for Controlling the Mutation Rate of an Evolutionary Algorithm19 Jun 2020 0 repositories listed
-
Parameterized MDPs and Reinforcement Learning Problems -- A Maximum Entropy Principle Based Framework17 Jun 2020 0 repositories listed
-
Q-learning with Logarithmic Regret16 Jun 2020 0 repositories listed
-
The Sample Complexity of Teaching-by-Reinforcement on Q-Learning16 Jun 2020 0 repositories listed
-
Runtime Adaptation in Wireless Sensor Nodes Using Structured Learning15 Jun 2020 0 repositories listed
-
Decorrelated Double Q-learning12 Jun 2020 0 repositories listed
-
Deep Reinforcement Learning for Neural Control12 Jun 2020 0 repositories listed
-
Human and Multi-Agent collaboration in a human-MARL teaming framework12 Jun 2020 0 repositories listed
-
Safety-guaranteed Reinforcement Learning based on Multi-class Support Vector Machine12 Jun 2020 0 repositories listed
-
12 Jun 2020 0 repositories listed Syntology 16 ran (of which 4 constructed an object rather than computing a result; 13 with no instrument failure: 0 honoured, 2 violated, 11 with no contract checked; 3 where Syntology's instrument failed) · 4 unverified (of 20 harvested samples) · 11 pointer-only (licence)
-
Exploration by Maximizing Rényi Entropy for Reward-Free RL Framework11 Jun 2020 0 repositories listed
-
Zeroth-Order Supervised Policy Improvement11 Jun 2020 0 repositories listed
-
Fitted Q-Learning for Relational Domains10 Jun 2020 0 repositories listed
-
Model-Free Algorithm and Regret Analysis for MDPs with Long-Term Constraints10 Jun 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning in a Realistic Limit Order Book Market Simulation10 Jun 2020 0 repositories listed
-
Privacy-Cost Management in Smart Meters with Mutual Information-Based Reinforcement Learning10 Jun 2020 0 repositories listed
-
Q-greedyUCB: a New Exploration Policy for Adaptive and Resource-efficient Scheduling10 Jun 2020 0 repositories listed
-
Self-Supervised Reinforcement Learning for Recommender Systems10 Jun 2020 0 repositories listed
-
Reinforcement Learning-Based Joint Self-Optimisation Method for the Fuzzy Logic Handover Algorithm in 5G HetNets9 Jun 2020 0 repositories listed
-
A Model-free Learning Algorithm for Infinite-horizon Average-reward MDPs with Near-optimal Regret8 Jun 2020 0 repositories listed
-
Balancing a CartPole System with Reinforcement Learning -- A Tutorial8 Jun 2020 0 repositories listed
-
Can Temporal-Difference and Q-Learning Learn Representation? A Mean-Field Theory8 Jun 2020 0 repositories listed
-
A Multi-step and Resilient Predictive Q-learning Algorithm for IoT with Human Operators in the Loop: A Case Study in Water Supply Networks6 Jun 2020 0 repositories listed
-
Logical Team Q-learning: An approach towards factored policies in cooperative MARL5 Jun 2020 0 repositories listed
-
Sample Complexity of Asynchronous Q-Learning: Sharper Analysis and Variance Reduction4 Jun 2020 0 repositories listed
-
Mitigating Bias in Face Recognition Using Skewness-Aware Reinforcement Learning1 Jun 2020 0 repositories listed
-
Hyperparameter optimization with REINFORCE and Transformers1 Jun 2020 0 repositories listed
-
Towards Understanding Cooperative Multi-Agent Q-Learning with Value Factorization31 May 2020 0 repositories listed
-
Learning-Based Joint User-AP Association and Resource Allocation in Ultra Dense Network28 May 2020 0 repositories listed
-
Active Measure Reinforcement Learning for Observation Cost Minimization26 May 2020 0 repositories listed
-
Deep Reinforcement Learning Based Power Allocation for D2D Network25 May 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.