Browse State-of-the-Art › Q-Learning › Papers, page 12
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 12 of 20: papers 1,101 to 1,200 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Neuromimetic Linear Systems -- Resilience and Learning10 May 2022 0 repositories listed
-
Vehicle management in a modular production context using Deep Q-Learning6 May 2022 0 repositories listed
-
Chemoreception and chemotaxis of a three-sphere swimmer5 May 2022 0 repositories listed
-
Q-Learning Scheduler for Multi Task Learning Through the use of Histogram of Task Uncertainty1 May 2022 0 repositories listed
-
Learning Value Functions from Undirected State-only Experience26 Apr 2022 0 repositories listed
-
Graph Neural Network based Agent in Google Research Football23 Apr 2022 0 repositories listed
-
Provably Efficient Kernelized Q-Learning21 Apr 2022 0 repositories listed
-
Joint Learning of Reward Machines and Policies in Environments with Partially Known Semantics20 Apr 2022 0 repositories listed
-
Efficient and practical quantum compiler towards multi-qubit systems with deep reinforcement learning14 Apr 2022 0 repositories listed
-
Optimizing the Long-Term Behaviour of Deep Reinforcement Learning for Pushing and Grasping7 Apr 2022 0 repositories listed
-
Q-learning with online random forests7 Apr 2022 0 repositories listed
-
Deep Q-learning of global optimizer of multiply model parameters for viscoelastic imaging1 Apr 2022 0 repositories listed
-
Functional Stability of Discounted Markov Decision Processes Using Economic MPC Dissipativity Theory31 Mar 2022 0 repositories listed
-
Neural Q-learning for solving PDEs31 Mar 2022 0 repositories listed
-
Investigating the Properties of Neural Network Representations in Reinforcement Learning30 Mar 2022 0 repositories listed
-
A Conservative Q-Learning approach for handling distribution shift in sepsis treatment strategies25 Mar 2022 0 repositories listed
-
The state-of-the-art review on resource allocation problem using artificial intelligence methods on various computing paradigms23 Mar 2022 0 repositories listed
-
A Note on Target Q-learning For Solving Finite MDPs with A Generative Oracle22 Mar 2022 0 repositories listed
-
Distributed Learning for Vehicular Dynamic Spectrum Access in Autonomous Driving22 Mar 2022 0 repositories listed
-
Infinite-Horizon Reach-Avoid Zero-Sum Games via Deep Reinforcement Learning18 Mar 2022 0 repositories listed
-
Reinforcement Learning for Optimal Control of a District Cooling Energy Plant14 Mar 2022 0 repositories listed
-
The Efficacy of Pessimism in Asynchronous Q-Learning14 Mar 2022 0 repositories listed
-
A Machine Learning Approach for Prosumer Management in Intraday Electricity Markets11 Mar 2022 0 repositories listed
-
Graph-based Reinforcement Learning meets Mixed Integer Programs: An application to 3D robot assembly discovery8 Mar 2022 0 repositories listed
-
Scalable multi-agent reinforcement learning for distributed control of residential energy flexibility7 Mar 2022 0 repositories listed
-
Offline Deep Reinforcement Learning for Dynamic Pricing of Consumer Credit6 Mar 2022 0 repositories listed
-
Target Network and Truncation Overcome The Deadly Triad in Q-Learning5 Mar 2022 0 repositories listed
-
A Learning Based Framework for Handling Uncertain Lead Times in Multi-Product Inventory Management2 Mar 2022 0 repositories listed
-
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise2 Mar 2022 0 repositories listed
-
Pessimistic Q-Learning for Offline Reinforcement Learning: Towards Optimal Sample Complexity28 Feb 2022 0 repositories listed
-
Whittle Index based Q-Learning for Wireless Edge Caching with Linear Function Approximation26 Feb 2022 0 repositories listed
-
Autonomous Warehouse Robot using Deep Q-Learning21 Feb 2022 0 repositories listed
-
PooL: Pheromone-inspired Communication Framework forLarge Scale Multi-Agent Reinforcement Learning20 Feb 2022 0 repositories listed
-
UAV Base Station Trajectory Optimization Based on Reinforcement Learning in Post-disaster Search and Rescue Operations17 Feb 2022 0 repositories listed
-
Artificial Intelligence and Auction Design12 Feb 2022 0 repositories listed
-
Regularized Q-learning11 Feb 2022 0 repositories listed
-
Intelligent Autonomous Intersection Management9 Feb 2022 0 repositories listed
-
Transferred Q-learning9 Feb 2022 0 repositories listed
-
Multiple Correlated Jammers Nullification using LSTM-based Deep Dueling Neural Network8 Feb 2022 0 repositories listed
-
Stochastic Gradient Descent with Dependent Data for Offline Reinforcement Learning6 Feb 2022 0 repositories listed
-
A deep Q-learning method for optimizing visual search strategies in backgrounds of dynamic noise28 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning with Spiking Q-learning21 Jan 2022 0 repositories listed
-
Optimal variance-reduced stochastic approximation in Banach spaces21 Jan 2022 0 repositories listed
-
A Family of Cognitively Realistic Parsing Environments for Deep Reinforcement Learning16 Jan 2022 0 repositories listed
-
Criticality-Based Varying Step-Number Algorithm for Reinforcement Learning13 Jan 2022 0 repositories listed
-
Task Independent Capsule-Based Agents for Deep Q-Learning11 Jan 2022 0 repositories listed
-
Age-of-information minimization via opportunistic sampling by an energy harvesting source8 Jan 2022 0 repositories listed
-
Sales Time Series Analytics Using Deep Q-Learning6 Jan 2022 0 repositories listed
-
Reinforcement Learning for Task Specifications with Action-Constraints2 Jan 2022 0 repositories listed
-
Operator Deep Q-Learning: Zero-Shot Reward Transferring in Reinforcement Learning1 Jan 2022 0 repositories listed
-
A Graph Attention Learning Approach to Antenna Tilt Optimization27 Dec 2021 0 repositories listed
-
Aerial Base Station Positioning and Power Control for Securing Communications: A Deep Q-Network Approach21 Dec 2021 0 repositories listed
-
Amortized Noisy Channel Neural Machine Translation16 Dec 2021 0 repositories listed
-
Finite-Sample Analysis of Decentralized Q-Learning for Stochastic Games15 Dec 2021 0 repositories listed
-
Teaching a Robot to Walk Using Reinforcement Learning13 Dec 2021 0 repositories listed
-
Control-Tutored Reinforcement Learning: Towards the Integration of Data-Driven and Model-Based Control11 Dec 2021 0 repositories listed
-
Quantum Architecture Search via Continual Reinforcement Learning10 Dec 2021 0 repositories listed
-
High-Dimensional Stock Portfolio Trading with Deep Reinforcement Learning9 Dec 2021 0 repositories listed
-
Application of Deep Reinforcement Learning to Payment Fraud8 Dec 2021 0 repositories listed
-
Convergence Results For Q-Learning With Experience Replay8 Dec 2021 0 repositories listed
-
Deep Q-Learning Market Makers in a Multi-Agent Simulated Stock Market8 Dec 2021 0 repositories listed
-
Replay For Safety8 Dec 2021 0 repositories listed
-
Pragmatic Implementation of Reinforcement Algorithms For Path Finding On Raspberry Pi7 Dec 2021 0 repositories listed
-
A Risk-Averse Preview-based Q-Learning Algorithm: Application to Highway Driving of Autonomous Vehicles6 Dec 2021 0 repositories listed
-
Faster Non-asymptotic Convergence for Double Q-learning1 Dec 2021 0 repositories listed
-
Finite Sample Analysis of Average-Reward TD Learning and Q-Learning1 Dec 2021 0 repositories listed
-
DeepCQ+: Robust and Scalable Routing with Multi-Agent Deep Reinforcement Learning for Highly Dynamic Networks29 Nov 2021 0 repositories listed
-
Final Adaptation Reinforcement Learning for N-Player Games29 Nov 2021 0 repositories listed
-
Count-Based Temperature Scheduling for Maximum Entropy Reinforcement Learning28 Nov 2021 0 repositories listed
-
Multicrew Scheduling and Routing in Road Network Restoration Based on Deep Q-learning24 Nov 2021 0 repositories listed
-
Reversible Action Design for Combinatorial Optimization with ReinforcementLearning24 Nov 2021 0 repositories listed
-
Multi-agent Bayesian Deep Reinforcement Learning for Microgrid Energy Management under Communication Failures22 Nov 2021 0 repositories listed
-
Aggressive Q-Learning with Ensembles: Achieving Both High Sample Efficiency and High Asymptotic Performance17 Nov 2021 0 repositories listed
-
Compressive Features in Offline Reinforcement Learning for Recommender Systems16 Nov 2021 0 repositories listed
-
Consecutive Task-oriented Dialog Policy Learning16 Nov 2021 0 repositories listed
-
Where to Look: A Unified Attention Model for Visual Recognition with Reinforcement Learning13 Nov 2021 0 repositories listed
-
Q-Learning for MDPs with General Spaces: Convergence and Near Optimality via Quantization under Weak Continuity12 Nov 2021 0 repositories listed
-
On Assessing The Safety of Reinforcement Learning algorithms Using Formal Methods8 Nov 2021 0 repositories listed
-
Supervised Advantage Actor-Critic for Recommender Systems5 Nov 2021 0 repositories listed
-
Towards Learning to Speak and Hear Through Multi-Agent Communication over a Continuous Acoustic Channel4 Nov 2021 0 repositories listed
-
Balanced Q-learning: Combining the Influence of Optimistic and Pessimistic Targets3 Nov 2021 0 repositories listed
-
2 Nov 2021 0 repositories listed
-
Decentralized Multi-Agent Reinforcement Learning: An Off-Policy Method31 Oct 2021 0 repositories listed
-
Throughput and Latency in the Distributed Q-Learning Random Access mMTC Networks30 Oct 2021 0 repositories listed
-
Learning to Communicate with Reinforcement Learning for an Adaptive Traffic Control System29 Oct 2021 0 repositories listed
-
Location-routing Optimisation for Urban Logistics Using Mobile Parcel Locker Based on Hybrid Q-Learning Algorithm29 Oct 2021 0 repositories listed
-
Cooperative Deep Q-learning Framework for Environments Providing Image Feedback28 Oct 2021 0 repositories listed
-
Temporal-Difference Value Estimation via Uncertainty-Guided Soft Updates28 Oct 2021 0 repositories listed
-
Finite Horizon Q-learning: Stability, Convergence, Simulations and an application on Smart Grids27 Oct 2021 0 repositories listed
-
V-Learning -- A Simple, Efficient, Decentralized Algorithm for Multiagent RL27 Oct 2021 0 repositories listed
-
Automating Control of Overestimation Bias for Reinforcement Learning26 Oct 2021 0 repositories listed
-
Can Q-Learning be Improved with Advice?25 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Simultaneous Sensing and Channel Access in Cognitive Networks24 Oct 2021 0 repositories listed
-
A Reinforcement Learning Approach to Parameter Selection for Distributed Optimal Power Flow22 Oct 2021 0 repositories listed
-
Can Q-learning solve Multi Armed Bantids?21 Oct 2021 0 repositories listed
-
A Q-Learning-based Approach for Distributed Beam Scheduling in mmWave Networks17 Oct 2021 0 repositories listed
-
Online Target Q-learning with Reverse Experience Replay: Efficiently finding the Optimal Policy for Linear MDPs16 Oct 2021 0 repositories listed
-
Value Penalized Q-Learning for Recommender Systems15 Oct 2021 0 repositories listed
-
Provably Efficient Multi-Agent Reinforcement Learning with Fully Decentralized Communication14 Oct 2021 0 repositories listed
-
On Improving Model-Free Algorithms for Decentralized Multi-Agent Reinforcement Learning12 Oct 2021 0 repositories listed