Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 126
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 126 of 152: papers 12,501 to 12,600 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A storage expansion planning framework using reinforcement learning and simulation-based optimization10 Jan 2020 0 repositories listed
-
Deep Interactive Reinforcement Learning for Path Following of Autonomous Underwater Vehicle10 Jan 2020 0 repositories listed
-
On Computation and Generalization of Generative Adversarial Imitation Learning9 Jan 2020 0 repositories listed
-
Reinforcement Learning Tracking Control for Robotic Manipulator With Kernel-Based Dynamic Model9 Jan 2020 0 repositories listed
-
EEG-based Drowsiness Estimation for Driving Safety using Deep Q-Learning8 Jan 2020 0 repositories listed
-
Multi-Agent Deep Reinforcement Learning for Cooperative Connected Vehicles8 Jan 2020 0 repositories listed
-
On Thompson Sampling for Smoother-than-Lipschitz Bandits8 Jan 2020 0 repositories listed
-
Perception and Navigation in Autonomous Systems in the Era of Learning: A Survey8 Jan 2020 0 repositories listed
-
Sample-based Distributional Policy Gradient8 Jan 2020 0 repositories listed
-
Decentralized Automotive Radar Spectrum Allocation to Avoid Mutual Interference Using Reinforcement Learning7 Jan 2020 0 repositories listed
-
Experimental Analysis of Reinforcement Learning Techniques for Spectrum Sharing Radar6 Jan 2020 0 repositories listed
-
High-speed Autonomous Drifting with Deep Reinforcement Learning6 Jan 2020 0 repositories listed
-
Generalizing Emergent Communication6 Jan 2020 0 repositories listed
-
Learning Reusable Options for Multi-Task Reinforcement Learning6 Jan 2020 0 repositories listed
-
Optimal Options for Multi-Task Reinforcement Learning Under Time Constraints6 Jan 2020 0 repositories listed
-
Universal Successor Features for Transfer Reinforcement Learning5 Jan 2020 0 repositories listed
-
Hierarchical Reinforcement Learning as a Model of Human Task Interleaving4 Jan 2020 0 repositories listed
-
Intelligent Roundabout Insertion using Deep Reinforcement Learning3 Jan 2020 0 repositories listed
-
Continuous-Discrete Reinforcement Learning for Hybrid Control in Robotics2 Jan 2020 0 repositories listed
-
Joint Goal and Strategy Inference across Heterogeneous Demonstrators via Reward Network Distillation2 Jan 2020 0 repositories listed
-
Zero-Shot Reinforcement Learning with Deep Attention Convolutional Neural Networks2 Jan 2020 0 repositories listed
-
A distributional view on multi objective policy optimization1 Jan 2020 0 repositories listed
-
A Game Theoretic Perspective on Model-Based Reinforcement Learning1 Jan 2020 0 repositories listed
-
Adaptive Droplet Routing in Digital Microfluidic Biochips Using Deep Reinforcement Learning1 Jan 2020 0 repositories listed
-
Batch Reinforcement Learning with Hyperparameter Gradients1 Jan 2020 0 repositories listed
-
Breaking the Curse of Many Agents: Provable Mean Embedding Q-Iteration for Mean-Field Reinforcement Learning1 Jan 2020 0 repositories listed
-
CoMic: Co-Training and Mimicry for Reusable Skills1 Jan 2020 0 repositories listed
-
Deep Randomized Least Squares Value Iteration1 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning with Implicit Human Feedback1 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning with Smooth Policy1 Jan 2020 0 repositories listed
-
Designing Optimal Dynamic Treatment Regimes: A Causal Reinforcement Learning Approach1 Jan 2020 0 repositories listed
-
Double Reinforcement Learning for Efficient and Robust Off-Policy Evaluation1 Jan 2020 0 repositories listed
-
Generative Adversarial Imitation Learning with Neural Network Parameterization: Global Optimality and Convergence Rate1 Jan 2020 0 repositories listed
-
Improving the Generalization of Visual Navigation Policies using Invariance Regularization1 Jan 2020 0 repositories listed
-
Inductive Bias-driven Reinforcement Learning For Efficient Schedules in Heterogeneous Clusters1 Jan 2020 0 repositories listed
-
Learning Fair Policies in Multi-Objective (Deep) Reinforcement Learning with Average and Discounted Rewards1 Jan 2020 0 repositories listed
-
Learning General-Purpose Controllers via Locally Communicating Sensorimotor Modules1 Jan 2020 0 repositories listed
-
Learning Representations in Reinforcement Learning: an Information Bottleneck Approach1 Jan 2020 0 repositories listed
-
Optimizing Multiagent Cooperation via Policy Evolution and Shared Experiences1 Jan 2020 0 repositories listed
-
OPtions as REsponses: Grounding behavioural hierarchies in multi-agent reinforcement learning1 Jan 2020 0 repositories listed
-
“Other-Play” for Zero-Shot Coordination1 Jan 2020 0 repositories listed
-
Reinforcement Learning with Differential Privacy1 Jan 2020 0 repositories listed
-
Reinforcement Learning with Goal-Distance Gradient1 Jan 2020 0 repositories listed
-
Responsive Safety in Reinforcement Learning1 Jan 2020 0 repositories listed
-
SVQN: Sequential Variational Soft Q-Learning Networks1 Jan 2020 0 repositories listed
-
The Natural Lottery Ticket Winner: Reinforcement Learning with Ordinary Neural Circuits1 Jan 2020 0 repositories listed
-
Way Off-Policy Batch Deep Reinforcement Learning of Human Preferences in Dialog1 Jan 2020 0 repositories listed
-
Information Theoretic Model Predictive Q-Learning31 Dec 2019 0 repositories listed
-
The Gambler's Problem and Beyond31 Dec 2019 0 repositories listed
-
Uncertainty-Based Out-of-Distribution Classification in Deep Reinforcement Learning31 Dec 2019 0 repositories listed
-
A New Framework for Query Efficient Active Imitation Learning30 Dec 2019 0 repositories listed
-
Deep Reinforced Self-Attention Masks for Abstractive Summarization (DR.SAS)30 Dec 2019 0 repositories listed
-
World Programs for Model-Based Learning and Planning in Compositional State and Action Spaces30 Dec 2019 0 repositories listed
-
Augmented Replay Memory in Reinforcement Learning With Continuous Control29 Dec 2019 0 repositories listed
-
Computational model discovery with reinforcement learning29 Dec 2019 0 repositories listed
-
Individual specialization in multi-task environments with multiagent reinforcement learners29 Dec 2019 0 repositories listed
-
Real-time Policy Distillation in Deep Reinforcement Learning29 Dec 2019 0 repositories listed
-
Speeding up reinforcement learning by combining attention and agency features29 Dec 2019 0 repositories listed
-
Crowdfunding Dynamics Tracking: A Reinforcement Learning Approach27 Dec 2019 0 repositories listed
-
Deep reinforcement learning for complex evaluation of one-loop diagrams in quantum field theory27 Dec 2019 0 repositories listed
-
Evolution Strategies Converges to Finite Differences27 Dec 2019 0 repositories listed
-
Quantum Logic Gate Synthesis as a Markov Decision Process27 Dec 2019 0 repositories listed
-
Quasi-Newton Trust Region Policy Optimization26 Dec 2019 0 repositories listed
-
Learning to Combat Compounding-Error in Model-Based Reinforcement Learning24 Dec 2019 0 repositories listed
-
A Survey of Deep Reinforcement Learning in Video Games23 Dec 2019 0 repositories listed
-
Direct and indirect reinforcement learning23 Dec 2019 0 repositories listed
-
Hamilton-Jacobi-Bellman Equations for Q-Learning in Continuous Time23 Dec 2019 0 repositories listed
-
Monte-Carlo Tree Search for Policy Optimization23 Dec 2019 0 repositories listed
-
Energy-Aware Multi-Server Mobile Edge Computing: A Deep Reinforcement Learning Approach22 Dec 2019 0 repositories listed
-
Online Reinforcement Learning of Optimal Threshold Policies for Markov Decision Processes21 Dec 2019 0 repositories listed
-
Predictive Coding for Boosting Deep Reinforcement Learning with Sparse Rewards21 Dec 2019 0 repositories listed
-
Exploiting the potential of deep reinforcement learning for classification tasks in high-dimensional and unstructured data20 Dec 2019 0 repositories listed
-
Mastering Complex Control in MOBA Games with Deep Reinforcement Learning20 Dec 2019 0 repositories listed
-
Optimizing Collision Avoidance in Dense Airspace using Deep Reinforcement Learning20 Dec 2019 0 repositories listed
-
Teaching robots to perceive time -- A reinforcement learning approach (Extended version)20 Dec 2019 0 repositories listed
-
Deep Reinforcement Learning Designed Shinnar-Le Roux RF Pulse using Root-Flipping: DeepRF_SLR19 Dec 2019 0 repositories listed
-
Deep Reinforcement Learning for Motion Planning of Mobile Robots19 Dec 2019 0 repositories listed
-
Deep Reinforcement Learning for Smart Home Energy Management19 Dec 2019 0 repositories listed
-
Extendable NFV-Integrated Control Method Using Reinforcement Learning19 Dec 2019 0 repositories listed
-
Analysing Deep Reinforcement Learning Agents Trained with Domain Randomisation18 Dec 2019 0 repositories listed
-
Benchmarking the Neural Linear Model for Regression18 Dec 2019 0 repositories listed
-
Learning to grow: control of material self-assembly using evolutionary reinforcement learning18 Dec 2019 0 repositories listed
-
Taming an autonomous surface vehicle for path following and collision avoidance using deep reinforcement learning18 Dec 2019 0 repositories listed
-
Unpaired Image Enhancement Featuring Reinforcement-Learning-Controlled Image Editing Software17 Dec 2019 0 repositories listed
-
Coordination in Adversarial Sequential Team Games via Multi-Agent Deep Reinforcement Learning16 Dec 2019 0 repositories listed
-
KARL: Knowledge-Aware Reasoning Memory Modeling with Reinforcement Learning of Vector Space16 Dec 2019 0 repositories listed
-
Planning with Abstract Learned Models While Learning Transferable Subtasks16 Dec 2019 0 repositories listed
-
Bayesian Linear Regression on Deep Representations14 Dec 2019 0 repositories listed
-
Fairness in Multi-agent Reinforcement Learning for Stock Trading14 Dec 2019 0 repositories listed
-
Natural Actor-Critic Converges Globally for Hierarchical Linear Quadratic Regulator14 Dec 2019 0 repositories listed
-
Resolving Congestions in the Air Traffic Management Domain via Multiagent Reinforcement Learning Methods14 Dec 2019 0 repositories listed
-
Spatial Influence-aware Reinforcement Learning for Intelligent Transportation System14 Dec 2019 0 repositories listed
-
Lessons from reinforcement learning for biological representations of space13 Dec 2019 0 repositories listed
-
More Efficient Off-Policy Evaluation through Regularized Targeted Learning13 Dec 2019 0 repositories listed
-
Provably Efficient Reinforcement Learning with Aggregated States13 Dec 2019 0 repositories listed
-
Recruitment-imitation Mechanism for Evolutionary Reinforcement Learning13 Dec 2019 0 repositories listed
-
Control-Tutored Reinforcement Learning12 Dec 2019 0 repositories listed
-
Improved Activity Forecasting for Generating Trajectories12 Dec 2019 0 repositories listed
-
Provably Efficient Exploration in Policy Optimization12 Dec 2019 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.