Browse State-of-the-Art › reinforcement-learning › Papers, page 135
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 135 of 135: papers 13,401 to 13,427 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Feature Construction for Inverse Reinforcement Learning1 Dec 2010 0 repositories listed
-
Interval Estimation for Reinforcement-Learning Algorithms in Continuous-State Domains1 Dec 2010 0 repositories listed
-
LSTD with Random Projections1 Dec 2010 0 repositories listed
-
Nonparametric Bayesian Policy Priors for Reinforcement Learning1 Dec 2010 0 repositories listed
-
PAC-Bayesian Model Selection for Reinforcement Learning1 Dec 2010 0 repositories listed
-
Predictive State Temporal Difference Learning1 Dec 2010 0 repositories listed
-
Fast Reinforcement Learning for Energy-Efficient Wireless Communications29 Sep 2010 0 repositories listed
-
Reinforcement Learning via AIXI Approximation13 Jul 2010 0 repositories listed
-
Computational Model of Music Sight Reading: A Reinforcement Learning Approach4 Jul 2010 0 repositories listed
-
Discrete MDL Predicts in Total Variation1 Dec 2009 0 repositories listed
-
Manifold Embeddings for Model-Based Reinforcement Learning under Partial Observability1 Dec 2009 0 repositories listed
-
Skill Discovery in Continuous Reinforcement Learning Domains using Skill Chaining1 Dec 2009 0 repositories listed
-
Solving Stochastic Games1 Dec 2009 0 repositories listed
-
Training Factor Graphs with Reinforcement Learning for Efficient MAP Inference1 Dec 2009 0 repositories listed
-
Hebbian Learning of Bayes Optimal Decisions1 Dec 2008 0 repositories listed
-
Multi-resolution Exploration in Continuous Spaces1 Dec 2008 0 repositories listed
-
Near-optimal Regret Bounds for Reinforcement Learning1 Dec 2008 0 repositories listed
-
Optimization on a Budget: A Reinforcement Learning Approach1 Dec 2008 0 repositories listed
-
Policy Search for Motor Primitives in Robotics1 Dec 2008 0 repositories listed
-
Regularized Policy Iteration1 Dec 2008 0 repositories listed
-
Structure Learning in Human Sequential Decision-Making1 Dec 2008 0 repositories listed
-
Temporal Difference Based Actor Critic Learning - Convergence and Neural Implementation1 Dec 2008 0 repositories listed
-
Fitted Q-iteration in continuous action-space MDPs1 Dec 2007 0 repositories listed
-
Online Linear Regression and Its Application to Model-Based Reinforcement Learning1 Dec 2007 0 repositories listed
-
Reinforcement Learning: A Survey1 May 1996 0 repositories listed
-
Accidental exploration through value predictors0 repositories listed
-
FROM DEEP LEARNING TO DEEP DEDUCING: AUTOMATICALLY TRACKING DOWN NASH EQUILIBRIUM THROUGH AUTONOMOUS NEURAL AGENT, A POSSIBLE MISSING STEP TOWARD GENERAL A.I.0 repositories listed