Browse State-of-the-Art › Q-Learning › Papers, page 9
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 9 of 20: papers 801 to 900 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Enhanced Q-Learning Approach to Finite-Time Reachability with Maximum Probability for Probabilistic Boolean Control Networks12 Dec 2023 0 repositories listed
-
Joint User Association, Interference Cancellation and Power Control for Multi-IRS Assisted UAV Communications8 Dec 2023 0 repositories listed
-
Two-Timescale Q-Learning with Function Approximation in Zero-Sum Stochastic Games8 Dec 2023 0 repositories listed
-
An efficient data-based off-policy Q-learning algorithm for optimal output feedback control of linear systems6 Dec 2023 0 repositories listed
-
A Q-learning approach to the continuous control problem of robot inverted pendulum balancing5 Dec 2023 0 repositories listed
-
Provable Reinforcement Learning for Networked Control Systems with Stochastic Packet Disordering5 Dec 2023 0 repositories listed
-
Algorithmic collusion under competitive design5 Dec 2023 0 repositories listed
-
Anomaly Detection via Learning-Based Sequential Controlled Sensing30 Nov 2023 0 repositories listed
-
Data-efficient Deep Reinforcement Learning for Vehicle Trajectory Control30 Nov 2023 0 repositories listed
-
OpenSense: An Open-World Sensing Framework for Incremental Learning and Dynamic Sensor Scheduling on Embedded Edge Devices29 Nov 2023 0 repositories listed
-
Q-learning Based Optimal False Data Injection Attack on Probabilistic Boolean Control Networks29 Nov 2023 0 repositories listed
-
Reinforcement Learning from Diffusion Feedback: Q* for Image Search27 Nov 2023 0 repositories listed
-
A Nearly Optimal and Low-Switching Algorithm for Reinforcement Learning with General Function Approximation26 Nov 2023 0 repositories listed
-
FRAC-Q-Learning: A Reinforcement Learning with Boredom Avoidance Processes for Social Robots26 Nov 2023 0 repositories listed
-
Projected Off-Policy Q-Learning (POP-QL) for Stabilizing Offline Reinforcement Learning25 Nov 2023 0 repositories listed
-
Approximation of Convex Envelope Using Reinforcement Learning24 Nov 2023 0 repositories listed
-
Efficient Open-world Reinforcement Learning via Knowledge Distillation and Autonomous Rule Discovery24 Nov 2023 0 repositories listed
-
Learning to Cooperate and Communicate Over Imperfect Channels24 Nov 2023 0 repositories listed
-
On optimal tracking portfolio in incomplete markets: The reinforcement learning approach24 Nov 2023 0 repositories listed
-
Machine learning-based decentralized TDMA for VLC IoT networks23 Nov 2023 0 repositories listed
-
Decentralised Q-Learning for Multi-Agent Markov Decision Processes with a Satisfiability Criterion21 Nov 2023 0 repositories listed
-
Offline Reinforcement Learning for Wireless Network Optimization with Mixture Datasets19 Nov 2023 0 repositories listed
-
Genetic Algorithm enhanced by Deep Reinforcement Learning in parent selection mechanism and mutation : Minimizing makespan in permutation flow shop scheduling problems10 Nov 2023 0 repositories listed
-
Advancing Algorithmic Trading: A Multi-Technique Enhancement of Deep Q-Network Models9 Nov 2023 0 repositories listed
-
Pointer Networks with Q-Learning for Combinatorial Optimization5 Nov 2023 0 repositories listed
-
Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments31 Oct 2023 0 repositories listed
-
DGFN: Double Generative Flow Networks30 Oct 2023 0 repositories listed
-
28 Oct 2023 0 repositories listed Syntology 5 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 2 where Syntology's instrument failed) · 5 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
Lifting the Veil: Unlocking the Power of Depth in Q-learning27 Oct 2023 0 repositories listed
-
Model-free Posterior Sampling via Learning Rate Randomization27 Oct 2023 0 repositories listed
-
Integrated Freeway Traffic Control Using Q-Learning with Adjacent Arterial Traffic Considerations25 Oct 2023 0 repositories listed
-
On the Convergence and Sample Complexity Analysis of Deep Q-Networks with ε-Greedy Exploration24 Oct 2023 0 repositories listed
-
Reinforcement learning based local path planning for mobile robot24 Oct 2023 0 repositories listed
-
AI on the Water: Applying DRL to Autonomous Vessel Navigation23 Oct 2023 0 repositories listed
-
Bad Values but Good Behavior: Learning Highly Misspecified Bandits and MDPs13 Oct 2023 0 repositories listed
-
Integrated Sensing and Communication Neighbor Discovery for MANET with Gossip Mechanism11 Oct 2023 0 repositories listed
-
Inverse Factorized Q-Learning for Cooperative Multi-agent Imitation Learning10 Oct 2023 0 repositories listed
-
Suppressing Overestimation in Q-Learning through Adversarial Behaviors10 Oct 2023 0 repositories listed
-
Dynamic value alignment through preference aggregation of multiple objectives9 Oct 2023 0 repositories listed
-
Diff-Transfer: Model-based Robotic Manipulation Skill Transfer via Differentiable Physics Simulation7 Oct 2023 0 repositories listed
-
Digital Twin Assisted Deep Reinforcement Learning for Online Admission Control in Sliced Network7 Oct 2023 0 repositories listed
-
Applying Reinforcement Learning to Option Pricing and Hedging6 Oct 2023 0 repositories listed
-
Optimal Control of District Cooling Energy Plant with Reinforcement Learning and MPC5 Oct 2023 0 repositories listed
-
A Deep Reinforcement Learning Approach for Interactive Search with Sentence-level Feedback3 Oct 2023 0 repositories listed
-
Finite-Time Analysis of Whittle Index based Q-Learning for Restless Multi-Armed Bandits with Neural Network Function Approximation3 Oct 2023 0 repositories listed
-
Using Reinforcement Learning to Optimize Responses in Care Processes: A Case Study on Aggression Incidents2 Oct 2023 0 repositories listed
-
Reinforcement learning adaptive fuzzy controller for lighting systems: application to aircraft cabin30 Sep 2023 0 repositories listed
-
Multi-Bellman operator for convergence of Q-learning with linear function approximation28 Sep 2023 0 repositories listed
-
Decoding trust: A reinforcement learning perspective26 Sep 2023 0 repositories listed
-
Adapting Double Q-Learning for Continuous Reinforcement Learning25 Sep 2023 0 repositories listed
-
UAV Swarm Deployment and Trajectory for 3D Area Coverage via Reinforcement Learning21 Sep 2023 0 repositories listed
-
Adaptive Multi-Agent Deep Reinforcement Learning for Timely Healthcare Interventions20 Sep 2023 0 repositories listed
-
Differentiable Quantum Architecture Search for Quantum Reinforcement Learning19 Sep 2023 0 repositories listed
-
Double Deep Q-Learning-based Path Selection and Service Placement for Latency-Sensitive Beyond 5G Applications18 Sep 2023 0 repositories listed
-
Q-Transformer: Scalable Offline Reinforcement Learning via Autoregressive Q-Functions18 Sep 2023 0 repositories listed
-
Self-Sustaining Multiple Access with Continual Deep Reinforcement Learning for Dynamic Metaverse Applications18 Sep 2023 0 repositories listed
-
Data-Driven H-infinity Control with a Real-Time and Efficient Reinforcement Learning Algorithm: An Application to Autonomous Mobility-on-Demand Systems16 Sep 2023 0 repositories listed
-
Harnessing Deep Q-Learning for Enhanced Statistical Arbitrage in High-Frequency Trading: A Comprehensive Exploration13 Sep 2023 0 repositories listed
-
A Q-learning Approach for Adherence-Aware Recommendations12 Sep 2023 0 repositories listed
-
Career Path Recommendations for Long-term Income Maximization: A Reinforcement Learning Approach11 Sep 2023 0 repositories listed
-
Convex Q Learning in a Stochastic Environment: Extended Version10 Sep 2023 0 repositories listed
-
Multi Agent DeepRL based Joint Power and Subchannel Allocation in IAB networks31 Aug 2023 0 repositories listed
-
Physics-Based Trajectory Design for Cellular-Connected UAV in Rainy Environments Based on Deep Reinforcement Learning31 Aug 2023 0 repositories listed
-
Actuator Trajectory Planning for UAVs with Overhead Manipulator using Reinforcement Learning24 Aug 2023 0 repositories listed
-
Improving Sample Efficiency of Model-Free Algorithms for Zero-Sum Markov Games17 Aug 2023 0 repositories listed
-
Reinforcement Learning for Battery Management in Dairy Farming17 Aug 2023 0 repositories listed
-
On-demand Cold Start Frequency Reduction with Off-Policy Reinforcement Learning in Serverless Computing15 Aug 2023 0 repositories listed
-
A Comparison of Classical and Deep Reinforcement Learning Methods for HVAC Control10 Aug 2023 0 repositories listed
-
Unsynchronized Decentralized Q-Learning: Two Timescale Analysis By Persistence7 Aug 2023 0 repositories listed
-
Deep Q-Network for Stochastic Process Environments7 Aug 2023 0 repositories listed
-
Minimax Optimal Q Learning with Nearest Neighbors3 Aug 2023 0 repositories listed
-
Stability of Multi-Agent Learning: Convergence in Network Games with Many Players26 Jul 2023 0 repositories listed
-
Adversarial Agents For Attacking Inaudible Voice Activated Devices23 Jul 2023 0 repositories listed
-
A Flexible Framework for Incorporating Patient Preferences Into Q-Learning22 Jul 2023 0 repositories listed
-
Distributed 3D-Beam Reforming for Hovering-Tolerant UAVs Communication over Coexistence: A Deep-Q Learning for Intelligent Space-Air-Ground Integrated Networks18 Jul 2023 0 repositories listed
-
Credit Assignment: Challenges and Opportunities in Developing Human-like AI Agents16 Jul 2023 0 repositories listed
-
Deep reinforcement learning for the dynamic vehicle dispatching problem: An event-based approach13 Jul 2023 0 repositories listed
-
Realtime Spectrum Monitoring via Reinforcement Learning -- A Comparison Between Q-Learning and Heuristic Methods11 Jul 2023 0 repositories listed
-
Investigating the Edge of Stability Phenomenon in Reinforcement Learning9 Jul 2023 0 repositories listed
-
The Value of Chess Squares8 Jul 2023 0 repositories listed
-
Offline Reinforcement Learning with Imbalanced Datasets6 Jul 2023 0 repositories listed
-
Elastic Decision Transformer5 Jul 2023 0 repositories listed
-
LLQL: Logistic Likelihood Q-Learning for Reinforcement Learning5 Jul 2023 0 repositories listed
-
Stability of Q-Learning Through Design and Optimism5 Jul 2023 0 repositories listed
-
Achieving Stable Training of Reinforcement Learning Agents in Bimodal Environments through Batch Learning3 Jul 2023 0 repositories listed
-
Is Risk-Sensitive Reinforcement Learning Properly Resolved?2 Jul 2023 0 repositories listed
-
Continuous-time q-learning for mean-field control problems28 Jun 2023 0 repositories listed
-
Evaluation of Reinforcement Learning Techniques for Trading on a Diverse Portfolio28 Jun 2023 0 repositories listed
-
Optimizing Credit Limit Adjustments Under Adversarial Goals Using Reinforcement Learning27 Jun 2023 0 repositories listed
-
RansomAI: AI-powered Ransomware for Stealthy Encryption27 Jun 2023 0 repositories listed
-
Decentralized Multi-Robot Formation Control Using Reinforcement Learning26 Jun 2023 0 repositories listed
-
Action Q-Transformer: Visual Explanation in Deep Reinforcement Learning with Encoder-Decoder Model using Action Query24 Jun 2023 0 repositories listed
-
Adaptive Ensemble Q-learning: Minimizing Estimation Bias via Error Feedback20 Jun 2023 0 repositories listed
-
Autonomous Driving with Deep Reinforcement Learning in CARLA Simulation20 Jun 2023 0 repositories listed
-
Vanishing Bias Heuristic-guided Reinforcement Learning Algorithm17 Jun 2023 0 repositories listed
-
Algorithmic Collusion in Auctions: Evidence from Controlled Laboratory Experiments15 Jun 2023 0 repositories listed
-
Residual Q-Learning: Offline and Online Policy Customization without Value15 Jun 2023 0 repositories listed
-
Privacy Risks in Reinforcement Learning for Household Robots15 Jun 2023 0 repositories listed
-
Model-based versus model-free feeding control and water quality monitoring for fish growth tracking in aquaculture systems14 Jun 2023 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.