Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 152
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 152 of 152: papers 15,101 to 15,113 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Hebbian Learning of Bayes Optimal Decisions1 Dec 2008 0 repositories listed
-
Multi-resolution Exploration in Continuous Spaces1 Dec 2008 0 repositories listed
-
Near-optimal Regret Bounds for Reinforcement Learning1 Dec 2008 0 repositories listed
-
Optimization on a Budget: A Reinforcement Learning Approach1 Dec 2008 0 repositories listed
-
Policy Search for Motor Primitives in Robotics1 Dec 2008 0 repositories listed
-
Regularized Policy Iteration1 Dec 2008 0 repositories listed
-
Structure Learning in Human Sequential Decision-Making1 Dec 2008 0 repositories listed
-
Temporal Difference Based Actor Critic Learning - Convergence and Neural Implementation1 Dec 2008 0 repositories listed
-
Fitted Q-iteration in continuous action-space MDPs1 Dec 2007 0 repositories listed
-
Online Linear Regression and Its Application to Model-Based Reinforcement Learning1 Dec 2007 0 repositories listed
-
Receding Horizon Differential Dynamic Programming1 Dec 2007 0 repositories listed
-
Accidental exploration through value predictors0 repositories listed
-
FROM DEEP LEARNING TO DEEP DEDUCING: AUTOMATICALLY TRACKING DOWN NASH EQUILIBRIUM THROUGH AUTONOMOUS NEURAL AGENT, A POSSIBLE MISSING STEP TOWARD GENERAL A.I.0 repositories listed