Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 122
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 122 of 152: papers 12,101 to 12,200 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
TOMA: Topological Map Abstraction for Reinforcement Learning11 May 2020 0 repositories listed
-
Accelerating Deep Neuroevolution on Distributed FPGAs for Reinforcement Learning Problems10 May 2020 0 repositories listed
-
An FPGA-Based On-Device Reinforcement Learning Approach using Online Sequential Learning10 May 2020 0 repositories listed
-
Optimal PID and Antiwindup Control Design as a Reinforcement Learning Problem10 May 2020 0 repositories listed
-
A Reinforcement Learning based approach for Multi-target Detection in Massive MIMO radar10 May 2020 0 repositories listed
-
Reinforcement Learning based Design of Linear Fixed Structure Controllers10 May 2020 0 repositories listed
-
Reinforcement Learning for Thermostatically Controlled Loads Control using Modelica and Python9 May 2020 0 repositories listed
-
Is Deep Reinforcement Learning Ready for Practical Applications in Healthcare? A Sensitivity Analysis of Duel-DDQN for Hemodynamic Management in Sepsis Patients8 May 2020 0 repositories listed
-
Synthesizing Safe Policies under Probabilistic Constraints with Reinforcement Learning and Bayesian Model Checking8 May 2020 0 repositories listed
-
Adaptive Dialog Policy Learning with Hindsight and User Modeling7 May 2020 0 repositories listed
-
Reinforcement Learning with Feedback Graphs7 May 2020 0 repositories listed
-
Robotic Arm Control and Task Training through Deep Reinforcement Learning6 May 2020 0 repositories listed
-
Safe Reinforcement Learning through Meta-learned Instincts6 May 2020 0 repositories listed
-
A Survey on Dialog Management: Recent Advances and Challenges5 May 2020 0 repositories listed
-
Generalized Planning With Deep Reinforcement Learning5 May 2020 0 repositories listed
-
Reinforcement Learning for UAV Autonomous Navigation, Mapping and Target Detection5 May 2020 0 repositories listed
-
Formal Policy Synthesis for Continuous-Space Systems via Reinforcement Learning4 May 2020 0 repositories listed
-
Generalized Reinforcement Meta Learning for Few-Shot Optimization4 May 2020 0 repositories listed
-
Hierarchical Decomposition of Nonlinear Dynamics and Control for System Identification and Policy Distillation4 May 2020 0 repositories listed
-
Multiagent Value Iteration Algorithms in Dynamic Programming and Reinforcement Learning4 May 2020 0 repositories listed
-
Noise Pollution in Hospital Readmission Prediction: Long Document Classification with Reinforcement Learning4 May 2020 0 repositories listed
-
Reward Constrained Interactive Recommendation with Natural Language Feedback4 May 2020 0 repositories listed
-
Setting up experimental Bell test with reinforcement learning4 May 2020 0 repositories listed
-
Multi-agent Reinforcement Learning for Decentralized Stable Matching3 May 2020 0 repositories listed
-
Deep Reinforcement Learning for Intelligent Transportation Systems: A Survey2 May 2020 0 repositories listed
-
Enhancing Text-based Reinforcement Learning Agents with Commonsense Knowledge2 May 2020 0 repositories listed
-
Optimal Beam Association for High Mobility mmWave Vehicular Networks: Lightweight Parallel Reinforcement Learning Approach2 May 2020 0 repositories listed
-
AMRL: Aggregated Memory For Reinforcement Learning1 May 2020 0 repositories listed
-
Episodic Reinforcement Learning with Associative Memory1 May 2020 0 repositories listed
-
Exploration in Reinforcement Learning with Deep Covering Options1 May 2020 0 repositories listed
-
Improving Robustness via Risk Averse Distributional Reinforcement Learning1 May 2020 0 repositories listed
-
Is Long Horizon Reinforcement Learning More Difficult Than Short Horizon Reinforcement Learning?1 May 2020 0 repositories listed
-
Keep Doing What Worked: Behavior Modelling Priors for Offline Reinforcement Learning1 May 2020 0 repositories listed
-
Learning Efficient Parameter Server Synchronization Policies for Distributed SGD1 May 2020 0 repositories listed
-
Learning Heuristics for Quantified Boolean Formulas through Reinforcement Learning1 May 2020 0 repositories listed
-
Learning the Arrow of Time for Problems in Reinforcement Learning1 May 2020 0 repositories listed
-
Model-based reinforcement learning for biological sequence design1 May 2020 0 repositories listed
-
Model Based Reinforcement Learning for Atari1 May 2020 0 repositories listed
-
Posterior sampling for multi-agent reinforcement learning: solving extensive games with imperfect information1 May 2020 0 repositories listed
-
Synthesizing Programmatic Policies that Inductively Generalize1 May 2020 0 repositories listed
-
The Ingredients of Real World Robotic Reinforcement Learning1 May 2020 0 repositories listed
-
Toward Evaluating Robustness of Deep Reinforcement Learning with Continuous Control1 May 2020 0 repositories listed
-
Bootstrap Latent-Predictive Representations for Multitask Reinforcement Learning30 Apr 2020 0 repositories listed
-
Breaking (Global) Barriers in Parallel Stochastic Optimization with Wait-Avoiding Group Averaging30 Apr 2020 0 repositories listed
-
Delay-aware Resource Allocation in Fog-assisted IoT Networks Through Reinforcement Learning30 Apr 2020 0 repositories listed
-
DSAC: Distributional Soft Actor Critic for Risk-Sensitive Reinforcement Learning30 Apr 2020 0 repositories listed
-
GCN-RL Circuit Designer: Transferable Transistor Sizing with Graph Neural Networks and Reinforcement Learning30 Apr 2020 0 repositories listed
-
Improving Factual Consistency Between a Response and Persona Facts30 Apr 2020 0 repositories listed
-
Out-of-the-box channel pruned networks30 Apr 2020 0 repositories listed
-
Plan-Space State Embeddings for Improved Reinforcement Learning30 Apr 2020 0 repositories listed
-
Reinforcement learning of minimalist grammars30 Apr 2020 0 repositories listed
-
Towards Embodied Scene Description30 Apr 2020 0 repositories listed
-
Unsupervised Learning of KB Queries in Task-Oriented Dialogs30 Apr 2020 0 repositories listed
-
Meta-Reinforcement Learning for Robotic Industrial Insertion Tasks29 Apr 2020 0 repositories listed
-
Molecular Design in Synthetically Accessible Chemical Space via Deep Reinforcement Learning29 Apr 2020 0 repositories listed
-
Reduced-Dimensional Reinforcement Learning Control using Singular Perturbation Approximations29 Apr 2020 0 repositories listed
-
Whittle index based Q-learning for restless bandits with average reward29 Apr 2020 0 repositories listed
-
Improving Sample Efficiency and Multi-Agent Communication in RL-based Train Rescheduling28 Apr 2020 0 repositories listed
-
The Immersion of Directed Multi-graphs in Embedding Fields. Generalisations28 Apr 2020 0 repositories listed
-
Adaptive model selection in photonic reservoir computing by reinforcement learning27 Apr 2020 0 repositories listed
-
Age-Aware Status Update Control for Energy Harvesting IoT Sensors via Reinforcement Learning27 Apr 2020 0 repositories listed
-
Can We Learn Heuristics For Graphical Model Inference Using Reinforcement Learning?27 Apr 2020 0 repositories listed
-
The Ingredients of Real-World Robotic Reinforcement Learning27 Apr 2020 0 repositories listed
-
A State Aggregation Approach for Solving Knapsack Problem with Deep Reinforcement Learning25 Apr 2020 0 repositories listed
-
Automatic low-bit hybrid quantization of neural networks through meta learning24 Apr 2020 0 repositories listed
-
PBCS : Efficient Exploration and Exploitation Using a Synergy between Reinforcement Learning and Motion Planning24 Apr 2020 0 repositories listed
-
Cooperative Perception with Deep Reinforcement Learning for Connected Vehicles23 Apr 2020 0 repositories listed
-
Divide-and-Conquer Monte Carlo Tree Search For Goal-Directed Planning23 Apr 2020 0 repositories listed
-
Guiding Robot Exploration in Reinforcement Learning via Automated Planning23 Apr 2020 0 repositories listed
-
Learning Dialog Policies from Weak Demonstrations23 Apr 2020 0 repositories listed
-
AutoEG: Automated Experience Grafting for Off-Policy Deep Reinforcement Learning22 Apr 2020 0 repositories listed
-
Flexible and Efficient Long-Range Planning Through Curious Exploration22 Apr 2020 0 repositories listed
-
Sequential Anomaly Detection using Inverse Reinforcement Learning22 Apr 2020 0 repositories listed
-
Almost Optimal Model-Free Reinforcement Learning via Reference-Advantage Decomposition21 Apr 2020 0 repositories listed
-
Never Stop Learning: The Effectiveness of Fine-Tuning in Robotic Reinforcement Learning21 Apr 2020 0 repositories listed
-
Reinforcement Learning to Optimize the Logistics Distribution Routes of Unmanned Aerial Vehicle21 Apr 2020 0 repositories listed
-
SIBRE: Self Improvement Based REwards for Adaptive Feedback in Reinforcement Learning21 Apr 2020 0 repositories listed
-
Attention Routing: track-assignment detailed routing using attention-based reinforcement learning20 Apr 2020 0 repositories listed
-
Data-Driven Learning and Load Ensemble Control20 Apr 2020 0 repositories listed
-
Learning as Reinforcement: Applying Principles of Neuroscience for More General Reinforcement Learning Agents20 Apr 2020 0 repositories listed
-
Tightening Exploration in Upper Confidence Reinforcement Learning20 Apr 2020 0 repositories listed
-
Variational Policy Propagation for Multi-agent Reinforcement Learning19 Apr 2020 0 repositories listed
-
Superkernel Neural Architecture Search for Image Denoising19 Apr 2020 0 repositories listed
-
Macro-Action-Based Deep Multi-Agent Reinforcement Learning18 Apr 2020 0 repositories listed
-
Modeling Survival in model-based Reinforcement Learning18 Apr 2020 0 repositories listed
-
Time Adaptive Reinforcement Learning18 Apr 2020 0 repositories listed
-
Approximate Inverse Reinforcement Learning from Vision-based Imitation Learning17 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning for Adaptive Learning Systems17 Apr 2020 0 repositories listed
-
F2A2: Flexible Fully-decentralized Approximate Actor-critic for Cooperative Multi-agent Reinforcement Learning17 Apr 2020 0 repositories listed
-
Goal-conditioned Batch Reinforcement Learning for Rotation Invariant Locomotion17 Apr 2020 0 repositories listed
-
Knowledge-guided Deep Reinforcement Learning for Interactive Recommendation17 Apr 2020 0 repositories listed
-
Show Us the Way: Learning to Manage Dialog from Demonstrations17 Apr 2020 0 repositories listed
-
Data-Driven Robust Control Using Reinforcement Learning16 Apr 2020 0 repositories listed
-
Reinforcement Learning for Safety-Critical Control under Model Uncertainty, using Control Lyapunov Functions and Control Barrier Functions16 Apr 2020 0 repositories listed
-
ActionSpotter: Deep Reinforcement Learning Framework for Temporal Action Spotting in Videos15 Apr 2020 0 repositories listed
-
Extending Deep Reinforcement Learning Frameworks in Cryptocurrency Market Making15 Apr 2020 0 repositories listed
-
Improving Input-Output Linearizing Controllers for Bipedal Robots via Reinforcement Learning15 Apr 2020 0 repositories listed
-
Safe deep reinforcement learning-based constrained optimal control scheme for active distribution networks15 Apr 2020 0 repositories listed
-
A Demonstration of Issues with Value-Based Multiobjective Reinforcement Learning Under Stochastic State Transitions14 Apr 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.