Browse State-of-the-Art › Reinforcement Learning › Papers, page 42
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 42 of 132: papers 4,101 to 4,200 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
6 Mar 2017 1 repository listed
-
3 Mar 2017 1 repository listed
-
3 Mar 2017 1 repository listed
-
3 Mar 2017 1 repository listed
-
2 Mar 2017 1 repository listed
-
1 Mar 2017 1 repository listed
-
28 Feb 2017 1 repository listed
-
27 Feb 2017 1 repository listed
-
22 Feb 2017 1 repository listed
-
21 Feb 2017 1 repository listed
-
21 Feb 2017 1 repository listed
-
21 Feb 2017 1 repository listed
-
20 Feb 2017 1 repository listed
-
19 Feb 2017 1 repository listed
-
8 Feb 2017 1 repository listed
-
25 Jan 2017 1 repository listed
-
16 Jan 2017 1 repository listed
-
15 Jan 2017 1 repository listed
-
10 Jan 2017 1 repository listed
-
9 Jan 2017 1 repository listed
-
A Survey of Deep Network Solutions for Learning Control in Robotics: From Reinforcement to Imitation21 Dec 2016 1 repository listed
-
19 Dec 2016 1 repository listed
-
1 Dec 2016 1 repository listed
-
1 Dec 2016 1 repository listed
-
26 Nov 2016 1 repository listed
-
26 Nov 2016 1 repository listed
-
25 Nov 2016 1 repository listed
-
22 Nov 2016 1 repository listed
-
13 Nov 2016 1 repository listed
-
11 Nov 2016 1 repository listed
-
11 Nov 2016 1 repository listed
-
7 Nov 2016 1 repository listed
-
5 Nov 2016 1 repository listed
-
1 Nov 2016 1 repository listed
-
13 Oct 2016 1 repository listed
-
6 Oct 2016 1 repository listed
-
3 Oct 2016 1 repository listed
-
27 Sep 2016 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 1 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
18 Sep 2016 1 repository listed
-
12 Sep 2016 1 repository listed
-
3 Sep 2016 1 repository listed
-
22 Aug 2016 1 repository listed
-
9 Aug 2016 1 repository listed
-
18 Jul 2016 1 repository listed
-
15 Jul 2016 1 repository listed
-
12 Jul 2016 1 repository listed
-
29 Jun 2016 1 repository listed
-
15 Jun 2016 1 repository listed Syntology official: harvested, nothing ran · 0 ran · 2 unverified (of 2 harvested samples)
-
12 Jun 2016 1 repository listed
-
8 Jun 2016 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
8 Jun 2016 1 repository listed
-
6 Jun 2016 1 repository listed
-
30 May 2016 1 repository listed
-
28 Apr 2016 1 repository listed
-
25 Mar 2016 1 repository listed
-
14 Mar 2016 1 repository listed
-
28 Feb 2016 1 repository listed
-
18 Jan 2016 1 repository listed
-
6 Jan 2016 1 repository listed
-
13 Dec 2015 1 repository listed
-
4 Dec 2015 1 repository listed
-
25 Nov 2015 1 repository listed
-
19 Nov 2015 1 repository listed
-
19 Nov 2015 1 repository listed
-
21 Sep 2015 1 repository listed
-
31 Jul 2015 1 repository listed
-
3 Jul 2015 1 repository listed
-
17 Jun 2015 1 repository listed
-
4 May 2015 1 repository listed
-
10 Feb 2015 1 repository listed
-
30 Apr 2014 1 repository listed
-
15 Apr 2014 1 repository listed
-
4 Apr 2014 1 repository listed
-
18 Feb 2014 1 repository listed
-
4 Feb 2014 1 repository listed
-
1 Dec 2013 1 repository listed
-
1 Dec 2011 1 repository listed
-
1 Jul 2008 1 repository listed
-
6 Aug 1999 1 repository listed
-
1 May 1992 1 repository listed
-
CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning18 Jul 2025 0 repositories listed
-
Aligning Humans and Robots via Reinforcement Learning from Implicit Human Feedback17 Jul 2025 0 repositories listed
-
Autonomous Resource Management in Microservice Systems via Reinforcement Learning17 Jul 2025 0 repositories listed
-
From Novelty to Imitation: Self-Distilled Rewards for Offline Reinforcement Learning17 Jul 2025 0 repositories listed
-
Spectral Bellman Method: Unifying Representation and Exploration in RL17 Jul 2025 0 repositories listed
-
VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning17 Jul 2025 0 repositories listed
-
A Survey of Explainable Reinforcement Learning: Targets, Methods and Needs16 Jul 2025 0 repositories listed
-
Distributional Reinforcement Learning on Path-dependent Options16 Jul 2025 0 repositories listed
-
Improving Reinforcement Learning Sample-Efficiency using Local Approximation16 Jul 2025 0 repositories listed
-
Scaling Up RL: Unlocking Diverse Reasoning in LLMs via Prolonged Training16 Jul 2025 0 repositories listed
-
Thought Purity: Defense Paradigm For Chain-of-Thought Attack16 Jul 2025 0 repositories listed
-
Functional Emotion Modeling in Biomimetic Reinforcement Learning15 Jul 2025 0 repositories listed
-
Local Pairwise Distance Matching for Backpropagation-Free Reinforcement Learning15 Jul 2025 0 repositories listed
-
Tactical Decision for Multi-UGV Confrontation with a Vision-Language Model-Based Commander15 Jul 2025 0 repositories listed
-
Continual Reinforcement Learning by Planning with Online World Models12 Jul 2025 0 repositories listed
-
CTRLS: Chain-of-Thought Reasoning via Latent State-Transition10 Jul 2025 0 repositories listed
-
From Curiosity to Competence: How World Models Interact with the Dynamics of Exploration10 Jul 2025 0 repositories listed
Syntology lines on 6 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.