Browse State-of-the-Art › Q-Learning › Papers, page 10
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 10 of 20: papers 901 to 1,000 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Pruning the Way to Reliable Policies: A Multi-Objective Deep Q-Learning Approach to Critical Care13 Jun 2023 0 repositories listed
-
Approximate information state based convergence analysis of recurrent Q-learning9 Jun 2023 0 repositories listed
-
Finite-Time Analysis of Minimax Q-Learning for Two-Player Zero-Sum Markov Games: Switching System Approach9 Jun 2023 0 repositories listed
-
Active Inference in Hebbian Learning Networks8 Jun 2023 0 repositories listed
-
Reinforcement Learning-Based Control of CrazyFlie 2.X Quadrotor6 Jun 2023 0 repositories listed
-
Deep Q-Learning versus Proximal Policy Optimization: Performance Comparison in a Material Sorting Task2 Jun 2023 0 repositories listed
-
IQL-TD-MPC: Implicit Q-Learning for Hierarchical Model Predictive Control1 Jun 2023 0 repositories listed
-
VA-learning as a more efficient alternative to Q-learning29 May 2023 0 repositories listed
-
Sample Complexity of Variance-reduced Distributionally Robust Q-learning28 May 2023 0 repositories listed
-
A Comparative Analysis of Portfolio Optimization Using Mean-Variance, Hierarchical Risk Parity, and Reinforcement Learning Approaches on the Indian Stock Market27 May 2023 0 repositories listed
-
Reinforcement Learning With Reward Machines in Stochastic Games27 May 2023 0 repositories listed
-
Sample Efficient Reinforcement Learning in Mixed Systems through Augmented Samples and Its Applications to Queueing Networks25 May 2023 0 repositories listed
-
RSRM: Reinforcement Symbolic Regression Machine24 May 2023 0 repositories listed
-
OER: Offline Experience Replay for Continual Offline Reinforcement Learning23 May 2023 0 repositories listed
-
A Framework for Provably Stable and Consistent Training of Deep Feedforward Networks20 May 2023 0 repositories listed
-
Bayesian Risk-Averse Q-Learning with Streaming Observations18 May 2023 0 repositories listed
-
The Blessing of Heterogeneity in Federated Q-Learning: Linear Speedup and Beyond18 May 2023 0 repositories listed
-
Model-Free Robust Average-Reward Reinforcement Learning17 May 2023 0 repositories listed
-
Smart Home Energy Management: VAE-GAN synthetic dataset generator and Q-learning14 May 2023 0 repositories listed
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm11 May 2023 0 repositories listed
-
Deep Q-Learning-based Distribution Network Reconfiguration for Reliability Improvement2 May 2023 0 repositories listed
-
BCQQ: Batch-Constraint Quantum Q-Learning with Cyclic Data Re-uploading27 Apr 2023 0 repositories listed
-
Safe Q-learning for continuous-time linear systems26 Apr 2023 0 repositories listed
-
Adaptive Services Function Chain Orchestration For Digital Health Twin Use Cases: Heuristic-boosted Q-Learning Approach25 Apr 2023 0 repositories listed
-
Learned Collusion25 Apr 2023 0 repositories listed
-
Graph Exploration for Effective Multi-agent Q-Learning19 Apr 2023 0 repositories listed
-
Quantum deep Q learning with distributed prioritized experience replay19 Apr 2023 0 repositories listed
-
A study on a Q-Learning algorithm application to a manufacturing assembly problem17 Apr 2023 0 repositories listed
-
Exploring the Noise Resilience of Successor Features and Predecessor Features Algorithms in One and Two-Dimensional Environments14 Apr 2023 0 repositories listed
-
Deep reinforcement learning applied to an assembly sequence planning problem with user preferences13 Apr 2023 0 repositories listed
-
Reinforcement Learning Based Minimum State-flipped Control for the Reachability of Boolean Control Networks11 Apr 2023 0 repositories listed
-
RELS-DQN: A Robust and Efficient Local Search Framework for Combinatorial Optimization11 Apr 2023 0 repositories listed
-
Deep Reinforcement Learning Based Optimal Infinite-Horizon Control of Probabilistic Boolean Control Networks7 Apr 2023 0 repositories listed
-
Full Gradient Deep Reinforcement Learning for Average-Reward Criterion7 Apr 2023 0 repositories listed
-
A Tutorial Introduction to Reinforcement Learning3 Apr 2023 0 repositories listed
-
Quantitative Trading using Deep Q Learning3 Apr 2023 0 repositories listed
-
Understanding Reinforcement Learning Algorithms: The Progress from Basic Q-learning to Proximal Policy Optimization31 Mar 2023 0 repositories listed
-
Q-Learning based system for path planning with unmanned aerial vehicles swarms in obstacle environments30 Mar 2023 0 repositories listed
-
Concentration of Contractive Stochastic Approximation: Additive and Multiplicative Noise28 Mar 2023 0 repositories listed
-
Distributed Multi-Agent Deep Q-Learning for Fast Roaming in IEEE 802.11ax Wi-Fi Systems25 Mar 2023 0 repositories listed
-
Specific investments under negotiated transfer pricing: effects of different surplus sharing parameters on managerial performance: An agent-based simulation with fuzzy Q-learning agents25 Mar 2023 0 repositories listed
-
Robust Path Following on Rivers Using Bootstrapped Reinforcement Learning24 Mar 2023 0 repositories listed
-
Artificial Intelligence and Dual Contract22 Mar 2023 0 repositories listed
-
Comparing NARS and Reinforcement Learning: An Analysis of ONA and Q-Learning Algorithms17 Mar 2023 0 repositories listed
-
Towards Real-World Applications of Personalized Anesthesia Using Policy Constraint Q Learning for Propofol Infusion Control17 Mar 2023 0 repositories listed
-
Self-Inspection Method of Unmanned Aerial Vehicles in Power Plants Using Deep Q-Network Reinforcement Learning16 Mar 2023 0 repositories listed
-
Smoothed Q-learning15 Mar 2023 0 repositories listed
-
The tree reconstruction game: phylogenetic reconstruction using reinforcement learning12 Mar 2023 0 repositories listed
-
Digital Twin-Assisted Knowledge Distillation Framework for Heterogeneous Federated Learning10 Mar 2023 0 repositories listed
-
Ignorance is Bliss: Robust Control via Information Gating10 Mar 2023 0 repositories listed
-
Learning Strategic Value and Cooperation in Multi-Player Stochastic Games through Side Payments9 Mar 2023 0 repositories listed
-
Environment Transformer and Policy Optimization for Model-Based Offline Reinforcement Learning7 Mar 2023 0 repositories listed
-
Exploration via Epistemic Value Estimation7 Mar 2023 0 repositories listed
-
Double A3C: Deep Reinforcement Learning on OpenAI Gym Games4 Mar 2023 0 repositories listed
-
Wasserstein Actor-Critic: Directed Exploration via Optimism for Continuous-Actions Control4 Mar 2023 0 repositories listed
-
Intelligent O-RAN Traffic Steering for URLLC Through Deep Reinforcement Learning3 Mar 2023 0 repositories listed
-
A Deep Reinforcement Learning Trader without Offline Training1 Mar 2023 0 repositories listed
-
Finite-sample Guarantees for Nash Q-learning with Linear Function Approximation1 Mar 2023 0 repositories listed
-
The Point to Which Soft Actor-Critic Converges1 Mar 2023 0 repositories listed
-
Minimizing the Outage Probability in a Markov Decision Process28 Feb 2023 0 repositories listed
-
A Finite Sample Complexity Bound for Distributionally Robust Q-learning26 Feb 2023 0 repositories listed
-
Q-Cogni: An Integrated Causal Reinforcement Learning Framework26 Feb 2023 0 repositories listed
-
On Bellman's principle of optimality and Reinforcement learning for safety-constrained Markov decision process25 Feb 2023 0 repositories listed
-
Gauss-Newton Temporal Difference Learning with Nonlinear Function Approximation25 Feb 2023 0 repositories listed
-
Kernel-Based Distributed Q-Learning: A Scalable Reinforcement Learning Approach for Dynamic Treatment Regimes21 Feb 2023 0 repositories listed
-
Robust Auto-landing Control of an agile Regional Jet Using Fuzzy Q-learning21 Feb 2023 0 repositories listed
-
Logit-Q Dynamics for Efficient Learning in Stochastic Teams20 Feb 2023 0 repositories listed
-
Forecasting and stabilizing chaotic regimes in two macroeconomic models via artificial intelligence technologies and control methods20 Feb 2023 0 repositories listed
-
Deep Offline Reinforcement Learning for Real-world Treatment Optimization Applications15 Feb 2023 0 repositories listed
-
Online Statistical Inference for Nonlinear Stochastic Approximation with Markovian Data15 Feb 2023 0 repositories listed
-
A Lifetime Extended Energy Management Strategy for Fuel Cell Hybrid Electric Vehicles via Self-Learning Fuzzy Reinforcement Learning13 Feb 2023 0 repositories listed
-
Computation Offloading for Uncertain Marine Tasks by Cooperation of UAVs and Vessels13 Feb 2023 0 repositories listed
-
Differentially Private Deep Q-Learning for Pattern Privacy Preservation in MEC Offloading9 Feb 2023 0 repositories listed
-
Catch Me If You Can: Improving Adversaries in Cyber-Security With Q-Learning Algorithms7 Feb 2023 0 repositories listed
-
MACOptions: Multi-Agent Learning with Centralized Controller and Options Framework7 Feb 2023 0 repositories listed
-
Offline Minimax Soft-Q-learning Under Realizability and Partial Coverage5 Feb 2023 0 repositories listed
-
Best Possible Q-Learning2 Feb 2023 0 repositories listed
-
Diversity Through Exclusion (DTE): Niche Identification for Reinforcement Learning through Value-Decomposition2 Feb 2023 0 repositories listed
-
Sample Complexity of Kernel-Based Q-Learning1 Feb 2023 0 repositories listed
-
Analyzing Robustness of the Deep Reinforcement Learning Algorithm in Ramp Metering Applications Considering False Data Injection Attack and Defense28 Jan 2023 0 repositories listed
-
RCsearcher: Reaction Center Identification in Retrosynthesis via Deep Q-Learning28 Jan 2023 0 repositories listed
-
The impact of surplus sharing on the outcomes of specific investments under negotiated transfer pricing: An agent-based simulation with fuzzy Q-learning agents28 Jan 2023 0 repositories listed
-
Single-Trajectory Distributionally Robust Reinforcement Learning27 Jan 2023 0 repositories listed
-
FedHQL: Federated Heterogeneous Q-Learning26 Jan 2023 0 repositories listed
-
Asymptotic Convergence and Performance of Multi-Agent Q-Learning Dynamics23 Jan 2023 0 repositories listed
-
Asynchronous Deep Double Duelling Q-Learning for Trading-Signal Execution in Limit Order Book Markets20 Jan 2023 0 repositories listed
-
Risk-Averse Reinforcement Learning via Dynamic Time-Consistent Risk Measures14 Jan 2023 0 repositories listed
-
Decentralized model-free reinforcement learning in stochastic games with average-reward objective13 Jan 2023 0 repositories listed
-
Hierarchical Deep Q-Learning Based Handover in Wireless Networks with Dual Connectivity13 Jan 2023 0 repositories listed
-
Multi-Power Level Q-Learning Algorithm for Random Access in NOMA mMTC Systems12 Jan 2023 0 repositories listed
-
Tuning Path Tracking Controllers for Autonomous Cars Using Reinforcement Learning9 Jan 2023 0 repositories listed
-
Contextual Conservative Q-Learning for Offline Reinforcement Learning3 Jan 2023 0 repositories listed
-
Deep Spectral Q-learning with Application to Mobile Health3 Jan 2023 0 repositories listed
-
Temporal Difference Learning with Compressed Updates: Error-Feedback meets Reinforcement Learning3 Jan 2023 0 repositories listed
-
Decoding surface codes with deep reinforcement learning and probabilistic policy reuse22 Dec 2022 0 repositories listed
-
Bandit approach to conflict-free multi-agent Q-learning in view of photonic implementation20 Dec 2022 0 repositories listed
-
Taming Lagrangian Chaos with Multi-Objective Reinforcement Learning19 Dec 2022 0 repositories listed
-
Offline Robot Reinforcement Learning with Uncertainty-Guided Human Expert Sampling16 Dec 2022 0 repositories listed
-
VOQL: Towards Optimal Regret in Model-free RL with Nonlinear Function Approximation12 Dec 2022 0 repositories listed
-
Frugal Reinforcement-based Active Learning9 Dec 2022 0 repositories listed