Browse State-of-the-Art › reinforcement-learning › Papers, page 110
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 110 of 135: papers 10,901 to 11,000 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Improving Sample Efficiency and Multi-Agent Communication in RL-based Train Rescheduling28 Apr 2020 0 repositories listed
-
Adaptive model selection in photonic reservoir computing by reinforcement learning27 Apr 2020 0 repositories listed
-
Age-Aware Status Update Control for Energy Harvesting IoT Sensors via Reinforcement Learning27 Apr 2020 0 repositories listed
-
Can We Learn Heuristics For Graphical Model Inference Using Reinforcement Learning?27 Apr 2020 0 repositories listed
-
The Ingredients of Real-World Robotic Reinforcement Learning27 Apr 2020 0 repositories listed
-
A State Aggregation Approach for Solving Knapsack Problem with Deep Reinforcement Learning25 Apr 2020 0 repositories listed
-
PBCS : Efficient Exploration and Exploitation Using a Synergy between Reinforcement Learning and Motion Planning24 Apr 2020 0 repositories listed
-
Cooperative Perception with Deep Reinforcement Learning for Connected Vehicles23 Apr 2020 0 repositories listed
-
Guiding Robot Exploration in Reinforcement Learning via Automated Planning23 Apr 2020 0 repositories listed
-
Learning Dialog Policies from Weak Demonstrations23 Apr 2020 0 repositories listed
-
AutoEG: Automated Experience Grafting for Off-Policy Deep Reinforcement Learning22 Apr 2020 0 repositories listed
-
Flexible and Efficient Long-Range Planning Through Curious Exploration22 Apr 2020 0 repositories listed
-
Sequential Anomaly Detection using Inverse Reinforcement Learning22 Apr 2020 0 repositories listed
-
Almost Optimal Model-Free Reinforcement Learning via Reference-Advantage Decomposition21 Apr 2020 0 repositories listed
-
Never Stop Learning: The Effectiveness of Fine-Tuning in Robotic Reinforcement Learning21 Apr 2020 0 repositories listed
-
Reinforcement Learning to Optimize the Logistics Distribution Routes of Unmanned Aerial Vehicle21 Apr 2020 0 repositories listed
-
SIBRE: Self Improvement Based REwards for Adaptive Feedback in Reinforcement Learning21 Apr 2020 0 repositories listed
-
Attention Routing: track-assignment detailed routing using attention-based reinforcement learning20 Apr 2020 0 repositories listed
-
Data-Driven Learning and Load Ensemble Control20 Apr 2020 0 repositories listed
-
Learning as Reinforcement: Applying Principles of Neuroscience for More General Reinforcement Learning Agents20 Apr 2020 0 repositories listed
-
Tightening Exploration in Upper Confidence Reinforcement Learning20 Apr 2020 0 repositories listed
-
Variational Policy Propagation for Multi-agent Reinforcement Learning19 Apr 2020 0 repositories listed
-
Superkernel Neural Architecture Search for Image Denoising19 Apr 2020 0 repositories listed
-
Macro-Action-Based Deep Multi-Agent Reinforcement Learning18 Apr 2020 0 repositories listed
-
Modeling Survival in model-based Reinforcement Learning18 Apr 2020 0 repositories listed
-
Time Adaptive Reinforcement Learning18 Apr 2020 0 repositories listed
-
Approximate Inverse Reinforcement Learning from Vision-based Imitation Learning17 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning for Adaptive Learning Systems17 Apr 2020 0 repositories listed
-
Goal-conditioned Batch Reinforcement Learning for Rotation Invariant Locomotion17 Apr 2020 0 repositories listed
-
Knowledge-guided Deep Reinforcement Learning for Interactive Recommendation17 Apr 2020 0 repositories listed
-
Show Us the Way: Learning to Manage Dialog from Demonstrations17 Apr 2020 0 repositories listed
-
Data-Driven Robust Control Using Reinforcement Learning16 Apr 2020 0 repositories listed
-
Reinforcement Learning for Safety-Critical Control under Model Uncertainty, using Control Lyapunov Functions and Control Barrier Functions16 Apr 2020 0 repositories listed
-
ActionSpotter: Deep Reinforcement Learning Framework for Temporal Action Spotting in Videos15 Apr 2020 0 repositories listed
-
Extending Deep Reinforcement Learning Frameworks in Cryptocurrency Market Making15 Apr 2020 0 repositories listed
-
Improving Input-Output Linearizing Controllers for Bipedal Robots via Reinforcement Learning15 Apr 2020 0 repositories listed
-
Safe deep reinforcement learning-based constrained optimal control scheme for active distribution networks15 Apr 2020 0 repositories listed
-
A Demonstration of Issues with Value-Based Multiobjective Reinforcement Learning Under Stochastic State Transitions14 Apr 2020 0 repositories listed
-
A reinforcement learning application of guided Monte Carlo Tree Search algorithm for beam orientation selection in radiation therapy14 Apr 2020 0 repositories listed
-
Actor-Critic Deep Reinforcement Learning for Solving Job Shop Scheduling Problems14 Apr 2020 0 repositories listed
-
Extrapolation in Gridworld Markov-Decision Processes14 Apr 2020 0 repositories listed
-
Reinforcement Learning Approach to Vibration Compensation for Dynamic Feed Drive Systems14 Apr 2020 0 repositories listed
-
A Deep Reinforcement Learning Framework for Continuous Intraday Market Bidding13 Apr 2020 0 repositories listed
-
A non-cooperative meta-modeling game for automated third-party calibrating, validating, and falsifying constitutive laws with parallelized adversarial attacks13 Apr 2020 0 repositories listed
-
Aspect and Opinion Aware Abstractive Review Summarization with Reinforced Hard Typed Decoder13 Apr 2020 0 repositories listed
-
Thinking While Moving: Deep Reinforcement Learning with Concurrent Control13 Apr 2020 0 repositories listed
-
Reinforcement Learning via Reasoning from Demonstration12 Apr 2020 0 repositories listed
-
Certifiable Robustness to Adversarial State Uncertainty in Deep Reinforcement Learning11 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning for Process Control: A Primer for Beginners11 Apr 2020 0 repositories listed
-
Reinforcement Learning via Gaussian Processes with Neural Network Dual Kernels10 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning (DRL): Another Perspective for Unsupervised Wireless Localization9 Apr 2020 0 repositories listed
-
Policy Gradient using Weak Derivatives for Reinforcement Learning9 Apr 2020 0 repositories listed
-
Re-conceptualising the Language Game Paradigm in the Framework of Multi-Agent Reinforcement Learning9 Apr 2020 0 repositories listed
-
Reinforced Anytime Bottom Up Rule Learning for Knowledge Graph Completion9 Apr 2020 0 repositories listed
-
Monte-Carlo Siamese Policy on Actor for Satellite Image Super Resolution8 Apr 2020 0 repositories listed
-
Resource Management for Blockchain-enabled Federated Learning: A Deep Reinforcement Learning Approach8 Apr 2020 0 repositories listed
-
Stochastic Approximation with Markov Noise: Analysis and applications in reinforcement learning8 Apr 2020 0 repositories listed
-
Online Constrained Model-based Reinforcement Learning7 Apr 2020 0 repositories listed
-
Networked Multi-Agent Reinforcement Learning with Emergent Communication6 Apr 2020 0 repositories listed
-
Technical Report: Adaptive Control for Linearizable Systems Using On-Policy Reinforcement Learning6 Apr 2020 0 repositories listed
-
Uniform State Abstraction For Reinforcement Learning6 Apr 2020 0 repositories listed
-
Weakly-Supervised Reinforcement Learning for Controllable Behavior6 Apr 2020 0 repositories listed
-
Multi-agent Reinforcement Learning for Resource Allocation in IoT networks with Edge Computing5 Apr 2020 0 repositories listed
-
Reinforcement Learning Architectures: SAC, TAC, and ESAC5 Apr 2020 0 repositories listed
-
Stylistic Dialogue Generation via Information-Guided Reinforcement Learning Strategy5 Apr 2020 0 repositories listed
-
A Deep Ensemble Multi-Agent Reinforcement Learning Approach for Air Traffic Control3 Apr 2020 0 repositories listed
-
Reinforcement Learning for Mixed-Integer Problems Based on MPC3 Apr 2020 0 repositories listed
-
Average Reward Adjusted Discounted Reinforcement Learning: Near-Blackwell-Optimal Policies for Real-World Applications2 Apr 2020 0 repositories listed
-
Continuous Motion Planning with Temporal Logic Specifications using Deep Neural Networks2 Apr 2020 0 repositories listed
-
Exploration of Reinforcement Learning for Event Camera using Car-like Robots2 Apr 2020 0 repositories listed
-
Safe Reinforcement Learning via Projection on a Safe Set: How to Achieve Optimality?2 Apr 2020 0 repositories listed
-
Value Driven Representation for Human-in-the-Loop Reinforcement Learning2 Apr 2020 0 repositories listed
-
Constrained-Space Optimization and Reinforcement Learning for Complex Tasks1 Apr 2020 0 repositories listed
-
Counterfactual Multi-Agent Reinforcement Learning with Graph Convolution Communication1 Apr 2020 0 repositories listed
-
Statistically Model Checking PCTL Specifications on Markov Decision Processes via Reinforcement Learning1 Apr 2020 0 repositories listed
-
Controlling Rayleigh-Bénard convection via Reinforcement Learning31 Mar 2020 0 repositories listed
-
Mimicking Evolution with Reinforcement Learning31 Mar 2020 0 repositories listed
-
Optimal Bidding Strategy without Exploration in Real-time Bidding31 Mar 2020 0 repositories listed
-
Robotic Table Tennis with Model-Free Reinforcement Learning31 Mar 2020 0 repositories listed
-
Model-Reference Reinforcement Learning Control of Autonomous Surface Vehicles with Uncertainties30 Mar 2020 0 repositories listed
-
Parallel Knowledge Transfer in Multi-Agent Reinforcement Learning29 Mar 2020 0 repositories listed
-
Learning medical triage from clinicians using Deep Q-Learning28 Mar 2020 0 repositories listed
-
A Distributional Analysis of Sampling-Based Reinforcement Learning Algorithms27 Mar 2020 0 repositories listed
-
Adaptive Reward-Poisoning Attacks against Reinforcement Learning27 Mar 2020 0 repositories listed
-
AirRL: A Reinforcement Learning Approach to Urban Air Quality Inference27 Mar 2020 0 repositories listed
-
Towards Better Opioid Antagonists Using Deep Reinforcement Learning26 Mar 2020 0 repositories listed
-
Black-box Off-policy Estimation for Infinite-Horizon Reinforcement Learning24 Mar 2020 0 repositories listed
-
Distributional Reinforcement Learning with Ensembles24 Mar 2020 0 repositories listed
-
Driver Modeling through Deep Reinforcement Learning and Behavioral Game Theory24 Mar 2020 0 repositories listed
-
Finite-Time Analysis of Stochastic Gradient Descent under Markov Randomness24 Mar 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning for Problems with Combined Individual and Team Reward24 Mar 2020 0 repositories listed
-
Q-Learning in Regularized Mean-field Games24 Mar 2020 0 repositories listed
-
Importance of using appropriate baselines for evaluation of data-efficiency in deep reinforcement learning for Atari23 Mar 2020 0 repositories listed
-
Incorporating Relational Background Knowledge into Reinforcement Learning via Differentiable Inductive Logic Programming23 Mar 2020 0 repositories listed
-
Learning to Walk: Spike Based Reinforcement Learning for Hexapod Robot Central Pattern Generation22 Mar 2020 0 repositories listed
-
Reinforcement Learning in Economics and Finance22 Mar 2020 0 repositories listed
-
Autonomous UAV Navigation: A DDPG-based Deep Reinforcement Learning Approach21 Mar 2020 0 repositories listed
-
Comprehensive Review of Deep Reinforcement Learning Methods and Applications in Economics21 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning with Robust and Smooth Policy21 Mar 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.