Browse State-of-the-Art › Q-Learning › Papers, page 5
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 5 of 20: papers 401 to 500 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
4 Jun 2019 1 repository listed Syntology official (archive's flag): 3 ran · 3 ran (of which 0 constructed an object rather than computing a result; 2 with no instrument failure: 2 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples)
-
25 May 2019 1 repository listed
-
24 May 2019 1 repository listed
-
24 May 2019 1 repository listed
-
20 May 2019 1 repository listed
-
16 May 2019 1 repository listed
-
15 May 2019 1 repository listed
-
6 May 2019 1 repository listed
-
1 May 2019 1 repository listed
-
23 Apr 2019 1 repository listed
-
14 Mar 2019 1 repository listed
-
11 Mar 2019 1 repository listed Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples) · 1 pointer-only (licence)
-
16 Feb 2019 1 repository listed
-
6 Feb 2019 1 repository listed
-
30 Jan 2019 1 repository listed
-
25 Jan 2019 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
24 Jan 2019 1 repository listed
-
22 Jan 2019 1 repository listed
-
13 Jan 2019 1 repository listed
-
18 Dec 2018 1 repository listed
-
21 Nov 2018 1 repository listed
-
19 Nov 2018 1 repository listed
-
13 Nov 2018 1 repository listed
-
1 Nov 2018 1 repository listed
-
22 Oct 2018 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples)
-
14 Oct 2018 1 repository listed
-
18 Sep 2018 1 repository listed
-
15 Sep 2018 1 repository listed
-
15 Sep 2018 1 repository listed Syntology 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
13 Sep 2018 1 repository listed
-
1 Aug 2018 1 repository listed
-
10 Jul 2018 1 repository listed
-
1 Jul 2018 1 repository listed
-
13 May 2018 1 repository listed
-
23 Apr 2018 1 repository listed
-
19 Apr 2018 1 repository listed
-
14 Apr 2018 1 repository listed
-
19 Mar 2018 1 repository listed
-
28 Feb 2018 1 repository listed
-
18 Feb 2018 1 repository listed
-
17 Feb 2018 1 repository listed
-
13 Dec 2017 1 repository listed
-
9 Dec 2017 1 repository listed
-
20 Nov 2017 1 repository listed
-
18 Oct 2017 1 repository listed
-
13 Sep 2017 1 repository listed
-
18 Aug 2017 1 repository listed
-
22 Jun 2017 1 repository listed
-
8 Jun 2017 1 repository listed
-
30 May 2017 1 repository listed
-
9 May 2017 1 repository listed
-
28 Feb 2017 1 repository listed
-
1 Dec 2016 1 repository listed
-
5 Nov 2016 1 repository listed
-
6 Oct 2016 1 repository listed
-
6 Jan 2016 1 repository listed
-
23 Nov 2015 1 repository listed
-
2 Jul 2015 1 repository listed
-
4 Dec 2003 1 repository listed
-
6 Aug 1999 1 repository listed
-
Evaluating Reinforcement Learning Algorithms for Navigation in Simulated Robotic Quadrupeds: A Comparative Study Inspired by Guide Dog Behaviour17 Jul 2025 0 repositories listed
-
A Data-Ensemble-Based Approach for Sample-Efficient LQ Control of Linear Time-Varying Systems30 Jun 2025 0 repositories listed
-
Reinforcement Learning-Based Policy Optimisation For Heterogeneous Radio Access18 Jun 2025 0 repositories listed
-
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning16 Jun 2025 0 repositories listed
-
ReinDSplit: Reinforced Dynamic Split Learning for Pest Recognition in Precision Agriculture16 Jun 2025 0 repositories listed
-
"What are my options?": Explaining RL Agents with Diverse Near-Optimal Alternatives (Extended)11 Jun 2025 0 repositories listed
-
Q-learning-based Hierarchical Cooperative Local Search for Steelmaking-continuous Casting Scheduling Problem10 Jun 2025 0 repositories listed
-
Regret-Optimal Q-Learning with Low Cost for Single-Agent and Federated Reinforcement Learning5 Jun 2025 0 repositories listed
-
Bridging the Performance Gap Between Target-Free and Target-Based Reinforcement Learning With Iterated Q-Learning4 Jun 2025 0 repositories listed
-
Improving Performance of Spike-based Deep Q-Learning using Ternary Neurons3 Jun 2025 0 repositories listed
-
Reinforcement Learning for Hanabi31 May 2025 0 repositories listed
-
Entropic Risk Optimization in Discounted MDPs: Sample Complexity Bounds with a Generative Model30 May 2025 0 repositories listed
-
On Global Convergence Rates for Federated Policy Gradient under Heterogeneous Environment29 May 2025 0 repositories listed
-
BOFormer: Learning to Solve Multi-Objective Bayesian Optimization via Non-Markovian RL28 May 2025 0 repositories listed
-
Learning to Charge More: A Theoretical Study of Collusion by Q-Learning Agents28 May 2025 0 repositories listed
-
A General-Purpose Theorem for High-Probability Bounds of Stochastic Approximation with Polyak Averaging27 May 2025 0 repositories listed
-
Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies22 May 2025 0 repositories listed
-
Reinforcement Learning for Stock Transactions22 May 2025 0 repositories listed
-
OPA-Pack: Object-Property-Aware Robotic Bin Packing19 May 2025 0 repositories listed
-
When a Reinforcement Learning Agent Encounters Unknown Unknowns19 May 2025 0 repositories listed
-
Imagination-Limited Q-Learning for Offline Reinforcement Learning18 May 2025 0 repositories listed
-
ShiQ: Bringing back Bellman to LLMs16 May 2025 0 repositories listed
-
Automatic Reward Shaping from Confounded Offline Data16 May 2025 0 repositories listed
-
Bias or Optimality? Disentangling Bayesian Inference and Learning Biases in Human Decision-Making12 May 2025 0 repositories listed
-
Convert Language Model into a Value-based Strategic Planner11 May 2025 0 repositories listed
-
A Large Language Model-Enhanced Q-learning for Capacitated Vehicle Routing Problem with Time Windows9 May 2025 0 repositories listed
-
Universal Approximation Theorem for Deep Q-Learning via FBSDE System9 May 2025 0 repositories listed
-
Merging and Disentangling Views in Visual Reinforcement Learning for Robotic Manipulation7 May 2025 0 repositories listed
-
VLM Q-Learning: Aligning Vision-Language Models for Interactive Decision-Making6 May 2025 0 repositories listed
-
Universal Approximation Theorem of Deep Q-Networks4 May 2025 0 repositories listed
-
Rank-One Modified Value Iteration3 May 2025 0 repositories listed
-
Dynamic and Distributed Routing in IoT Networks based on Multi-Objective Q-Learning1 May 2025 0 repositories listed
-
Learning Neural Control Barrier Functions from Offline Data with Conservatism1 May 2025 0 repositories listed
-
Q-Learning with Clustered-SMART (cSMART) Data: Examining Moderators in the Construction of Clustered Adaptive Interventions1 May 2025 0 repositories listed
-
Interactive Double Deep Q-network: Integrating Human Interventions and Evaluative Predictions in Reinforcement Learning of Autonomous Driving28 Apr 2025 0 repositories listed
-
Non-Asymptotic Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes25 Apr 2025 0 repositories listed
-
SAPO-RL: Sequential Actuator Placement Optimization for Fuselage Assembly via Reinforcement Learning24 Apr 2025 0 repositories listed
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.