Browse State-of-the-Art › Q-Learning › Papers, page 20
Q-Learning
Papers archive 2025-07-28
archive papers tagged: 1,918 · with a code link: 463 · where Syntology ran a sample: 119 (102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (119 of 1,918 tagged: 102 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 20 of 20: papers 1,901 to 1,918 of 1,918, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Q-learning for Optimal Control of Continuous-time Systems11 Oct 2014 0 repositories listed
-
Learning to Cooperate via Policy Search7 Aug 2014 0 repositories listed
-
Reinforcement Learning Based Algorithm for the Maximization of EV Charging Station Revenue4 Jul 2014 0 repositories listed
-
Personalized Medical Treatments Using Novel Reinforcement Learning Algorithms16 Jun 2014 0 repositories listed
-
Single-Agent vs. Multi-Agent Techniques for Concurrent Reinforcement Learning of Negotiation Dialogue Policies1 Jun 2014 0 repositories listed
-
Empirically Evaluating Multiagent Learning Algorithms31 Jan 2014 0 repositories listed
-
Adaptive Stochastic Resource Control: A Machine Learning Approach15 Jan 2014 0 repositories listed
-
Optimal Demand Response Using Device Based Reinforcement Learning8 Jan 2014 0 repositories listed
-
Two Timescale Convergent Q-learning for Sleep--Scheduling in Wireless Sensor Networks27 Dec 2013 0 repositories listed
-
Q-learning optimization in a multi-agents system for image segmentation23 Nov 2013 0 repositories listed
-
Risk-sensitive Reinforcement Learning8 Nov 2013 0 repositories listed
-
Approximate Kalman Filter Q-Learning for Continuous State-Space MDPs26 Sep 2013 0 repositories listed
-
The association problem in wireless networks: a Policy Gradient Reinforcement Learning approach11 Jun 2013 0 repositories listed
-
Projective simulation for classical learning agents: a comprehensive investigation7 May 2013 0 repositories listed
-
Hybrid Q-Learning Applied to Ubiquitous recommender system10 Mar 2013 0 repositories listed
-
Speedy Q-Learning1 Dec 2011 0 repositories listed
-
1 Dec 2010 0 repositories listed
-
Convergent Temporal-Difference Learning with Arbitrary Smooth Function Approximation1 Dec 2009 0 repositories listed