Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 123
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 123 of 152: papers 12,201 to 12,300 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A reinforcement learning application of guided Monte Carlo Tree Search algorithm for beam orientation selection in radiation therapy14 Apr 2020 0 repositories listed
-
Actor-Critic Deep Reinforcement Learning for Solving Job Shop Scheduling Problems14 Apr 2020 0 repositories listed
-
Extrapolation in Gridworld Markov-Decision Processes14 Apr 2020 0 repositories listed
-
Reinforcement Learning Approach to Vibration Compensation for Dynamic Feed Drive Systems14 Apr 2020 0 repositories listed
-
A Deep Reinforcement Learning Framework for Continuous Intraday Market Bidding13 Apr 2020 0 repositories listed
-
A non-cooperative meta-modeling game for automated third-party calibrating, validating, and falsifying constitutive laws with parallelized adversarial attacks13 Apr 2020 0 repositories listed
-
Aspect and Opinion Aware Abstractive Review Summarization with Reinforced Hard Typed Decoder13 Apr 2020 0 repositories listed
-
K-spin Hamiltonian for quantum-resolvable Markov decision processes13 Apr 2020 0 repositories listed
-
Thinking While Moving: Deep Reinforcement Learning with Concurrent Control13 Apr 2020 0 repositories listed
-
Reinforcement Learning via Reasoning from Demonstration12 Apr 2020 0 repositories listed
-
Certifiable Robustness to Adversarial State Uncertainty in Deep Reinforcement Learning11 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning for Process Control: A Primer for Beginners11 Apr 2020 0 repositories listed
-
Reinforcement Learning via Gaussian Processes with Neural Network Dual Kernels10 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning (DRL): Another Perspective for Unsupervised Wireless Localization9 Apr 2020 0 repositories listed
-
Policy Gradient using Weak Derivatives for Reinforcement Learning9 Apr 2020 0 repositories listed
-
Quantifying the Impact of Non-Stationarity in Reinforcement Learning-Based Traffic Signal Control9 Apr 2020 0 repositories listed
-
Re-conceptualising the Language Game Paradigm in the Framework of Multi-Agent Reinforcement Learning9 Apr 2020 0 repositories listed
-
Reinforced Anytime Bottom Up Rule Learning for Knowledge Graph Completion9 Apr 2020 0 repositories listed
-
Adaptive Stress Testing without Domain Heuristics using Go-Explore8 Apr 2020 0 repositories listed
-
Monte-Carlo Siamese Policy on Actor for Satellite Image Super Resolution8 Apr 2020 0 repositories listed
-
Resource Management for Blockchain-enabled Federated Learning: A Deep Reinforcement Learning Approach8 Apr 2020 0 repositories listed
-
Stochastic Approximation with Markov Noise: Analysis and applications in reinforcement learning8 Apr 2020 0 repositories listed
-
Online Constrained Model-based Reinforcement Learning7 Apr 2020 0 repositories listed
-
Optimistic Agent: Accurate Graph-Based Value Estimation for More Successful Visual Navigation7 Apr 2020 0 repositories listed
-
Intrinsic Exploration as Multi-Objective RL6 Apr 2020 0 repositories listed
-
Networked Multi-Agent Reinforcement Learning with Emergent Communication6 Apr 2020 0 repositories listed
-
Technical Report: Adaptive Control for Linearizable Systems Using On-Policy Reinforcement Learning6 Apr 2020 0 repositories listed
-
Uniform State Abstraction For Reinforcement Learning6 Apr 2020 0 repositories listed
-
Weakly-Supervised Reinforcement Learning for Controllable Behavior6 Apr 2020 0 repositories listed
-
Multi-agent Reinforcement Learning for Resource Allocation in IoT networks with Edge Computing5 Apr 2020 0 repositories listed
-
Reinforced Multi-task Approach for Multi-hop Question Generation5 Apr 2020 0 repositories listed
-
Reinforcement Learning Architectures: SAC, TAC, and ESAC5 Apr 2020 0 repositories listed
-
Stylistic Dialogue Generation via Information-Guided Reinforcement Learning Strategy5 Apr 2020 0 repositories listed
-
A Deep Ensemble Multi-Agent Reinforcement Learning Approach for Air Traffic Control3 Apr 2020 0 repositories listed
-
Reinforcement Learning for Mixed-Integer Problems Based on MPC3 Apr 2020 0 repositories listed
-
Average Reward Adjusted Discounted Reinforcement Learning: Near-Blackwell-Optimal Policies for Real-World Applications2 Apr 2020 0 repositories listed
-
Continuous Motion Planning with Temporal Logic Specifications using Deep Neural Networks2 Apr 2020 0 repositories listed
-
Exploration of Reinforcement Learning for Event Camera using Car-like Robots2 Apr 2020 0 repositories listed
-
Safe Reinforcement Learning via Projection on a Safe Set: How to Achieve Optimality?2 Apr 2020 0 repositories listed
-
Value Driven Representation for Human-in-the-Loop Reinforcement Learning2 Apr 2020 0 repositories listed
-
Constrained-Space Optimization and Reinforcement Learning for Complex Tasks1 Apr 2020 0 repositories listed
-
Counterfactual Multi-Agent Reinforcement Learning with Graph Convolution Communication1 Apr 2020 0 repositories listed
-
Statistically Model Checking PCTL Specifications on Markov Decision Processes via Reinforcement Learning1 Apr 2020 0 repositories listed
-
Controlling Rayleigh-Bénard convection via Reinforcement Learning31 Mar 2020 0 repositories listed
-
Leverage the Average: an Analysis of KL Regularization in RL31 Mar 2020 0 repositories listed
-
Mimicking Evolution with Reinforcement Learning31 Mar 2020 0 repositories listed
-
Optimal Bidding Strategy without Exploration in Real-time Bidding31 Mar 2020 0 repositories listed
-
Robotic Table Tennis with Model-Free Reinforcement Learning31 Mar 2020 0 repositories listed
-
Model-Reference Reinforcement Learning Control of Autonomous Surface Vehicles with Uncertainties30 Mar 2020 0 repositories listed
-
Parallel Knowledge Transfer in Multi-Agent Reinforcement Learning29 Mar 2020 0 repositories listed
-
When Autonomous Systems Meet Accuracy and Transferability through AI: A Survey29 Mar 2020 0 repositories listed
-
Learning medical triage from clinicians using Deep Q-Learning28 Mar 2020 0 repositories listed
-
A Distributional Analysis of Sampling-Based Reinforcement Learning Algorithms27 Mar 2020 0 repositories listed
-
Adaptive Reward-Poisoning Attacks against Reinforcement Learning27 Mar 2020 0 repositories listed
-
AirRL: A Reinforcement Learning Approach to Urban Air Quality Inference27 Mar 2020 0 repositories listed
-
Towards Better Opioid Antagonists Using Deep Reinforcement Learning26 Mar 2020 0 repositories listed
-
ACNMP: Skill Transfer and Task Extrapolation through Learning from Demonstration and Reinforcement Learning via Representation Sharing25 Mar 2020 0 repositories listed
-
Black-box Off-policy Estimation for Infinite-Horizon Reinforcement Learning24 Mar 2020 0 repositories listed
-
Distributional Reinforcement Learning with Ensembles24 Mar 2020 0 repositories listed
-
Driver Modeling through Deep Reinforcement Learning and Behavioral Game Theory24 Mar 2020 0 repositories listed
-
Finite-Time Analysis of Stochastic Gradient Descent under Markov Randomness24 Mar 2020 0 repositories listed
-
Learning Compact Reward for Image Captioning24 Mar 2020 0 repositories listed
-
Learning to Play Soccer by Reinforcement and Applying Sim-to-Real to Compete in the Real World24 Mar 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning for Problems with Combined Individual and Team Reward24 Mar 2020 0 repositories listed
-
Q-Learning in Regularized Mean-field Games24 Mar 2020 0 repositories listed
-
Importance of using appropriate baselines for evaluation of data-efficiency in deep reinforcement learning for Atari23 Mar 2020 0 repositories listed
-
Incorporating Relational Background Knowledge into Reinforcement Learning via Differentiable Inductive Logic Programming23 Mar 2020 0 repositories listed
-
Learning to Walk: Spike Based Reinforcement Learning for Hexapod Robot Central Pattern Generation22 Mar 2020 0 repositories listed
-
Reinforcement Learning in Economics and Finance22 Mar 2020 0 repositories listed
-
Autonomous UAV Navigation: A DDPG-based Deep Reinforcement Learning Approach21 Mar 2020 0 repositories listed
-
Comprehensive Review of Deep Reinforcement Learning Methods and Applications in Economics21 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning with Robust and Smooth Policy21 Mar 2020 0 repositories listed
-
Distributed Reinforcement Learning for Cooperative Multi-Robot Object Manipulation21 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning with Weighted Q-Learning20 Mar 2020 0 repositories listed
-
Deep Sets for Generalization in RL20 Mar 2020 0 repositories listed
-
Deep Constrained Q-learning20 Mar 2020 0 repositories listed
-
Exchangeable Input Representations for Reinforcement Learning19 Mar 2020 0 repositories listed
-
Reinforcement learning enabled cooperative spectrum sensing in cognitive radio networks19 Mar 2020 0 repositories listed
-
Towards Cognitive Routing based on Deep Reinforcement Learning19 Mar 2020 0 repositories listed
-
Generating Socially Acceptable Perturbations for Efficient Evaluation of Autonomous Vehicles18 Mar 2020 0 repositories listed
-
Placement Optimization with Deep Reinforcement Learning18 Mar 2020 0 repositories listed
-
Viewport-Aware Deep Reinforcement Learning Approach for 360ᵒ Video Caching18 Mar 2020 0 repositories listed
-
Stop-and-Go: Exploring Backdoor Attacks on Deep Reinforcement Learning-based Traffic Congestion Control Systems17 Mar 2020 0 repositories listed
-
Improving Performance in Reinforcement Learning by Breaking Generalization in Neural Networks16 Mar 2020 0 repositories listed
-
Reinforcement Learning for Electricity Network Operation16 Mar 2020 0 repositories listed
-
Model-based Reinforcement Learning for Decentralized Multiagent Rendezvous15 Mar 2020 0 repositories listed
-
A General Framework for Learning Mean-Field Games13 Mar 2020 0 repositories listed
-
Optimizing Medical Treatment for Sepsis in Intensive Care: from Reinforcement Learning to Pre-Trial Evaluation13 Mar 2020 0 repositories listed
-
13 Mar 2020 0 repositories listed
-
Analyzing Visual Representations in Embodied Navigation Tasks12 Mar 2020 0 repositories listed
-
Heterogeneous Relational Reasoning in Knowledge Graphs with Reinforcement Learning12 Mar 2020 0 repositories listed
-
Automatic Curriculum Learning For Deep RL: A Short Survey10 Mar 2020 0 repositories listed
-
Curriculum Learning for Reinforcement Learning Domains: A Framework and Survey10 Mar 2020 0 repositories listed
-
Mobility Management for Cellular-Connected UAVs: A Learning-Based Approach10 Mar 2020 0 repositories listed
-
Privacy-Cost Management in Smart Meters Using Deep Reinforcement Learning10 Mar 2020 0 repositories listed
-
Reinforcement Learning for Mitigating Intermittent Interference in Terahertz Communication Networks10 Mar 2020 0 repositories listed
-
SQUIRL: Robust and Efficient Learning from Video Demonstration of Long-Horizon Robotic Manipulation Tasks10 Mar 2020 0 repositories listed
-
10 Mar 2020 0 repositories listed Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Advancing Renewable Electricity Consumption With Reinforcement Learning9 Mar 2020 0 repositories listed
-
Human AI interaction loop training: New approach for interactive reinforcement learning9 Mar 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.