Browse State-of-the-Art › Reinforcement Learning › Papers, page 95
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 95 of 132: papers 9,401 to 9,500 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Data-Driven Robust Control Using Reinforcement Learning16 Apr 2020 0 repositories listed
-
Order Matters: Generating Progressive Explanations for Planning Tasks in Human-Robot Teaming16 Apr 2020 0 repositories listed
-
Reinforcement Learning for Safety-Critical Control under Model Uncertainty, using Control Lyapunov Functions and Control Barrier Functions16 Apr 2020 0 repositories listed
-
ActionSpotter: Deep Reinforcement Learning Framework for Temporal Action Spotting in Videos15 Apr 2020 0 repositories listed
-
Bootstrapped model learning and error correction for planning with uncertainty in model-based RL15 Apr 2020 0 repositories listed
-
Contextual-Bandit Anomaly Detection for IoT Data in Distributed Hierarchical Edge Computing15 Apr 2020 0 repositories listed
-
Extending Deep Reinforcement Learning Frameworks in Cryptocurrency Market Making15 Apr 2020 0 repositories listed
-
Improving Input-Output Linearizing Controllers for Bipedal Robots via Reinforcement Learning15 Apr 2020 0 repositories listed
-
Joint User Pairing and Association for Multicell NOMA: A Pointer Network-based Approach15 Apr 2020 0 repositories listed
-
lamBERT: Language and Action Learning Using Multimodal BERT15 Apr 2020 0 repositories listed
-
Safe deep reinforcement learning-based constrained optimal control scheme for active distribution networks15 Apr 2020 0 repositories listed
-
A Demonstration of Issues with Value-Based Multiobjective Reinforcement Learning Under Stochastic State Transitions14 Apr 2020 0 repositories listed
-
A reinforcement learning application of guided Monte Carlo Tree Search algorithm for beam orientation selection in radiation therapy14 Apr 2020 0 repositories listed
-
Actor-Critic Deep Reinforcement Learning for Solving Job Shop Scheduling Problems14 Apr 2020 0 repositories listed
-
Adversarial Evaluation of Autonomous Vehicles in Lane-Change Scenarios14 Apr 2020 0 repositories listed
-
Extrapolation in Gridworld Markov-Decision Processes14 Apr 2020 0 repositories listed
-
Reinforcement Learning Approach to Vibration Compensation for Dynamic Feed Drive Systems14 Apr 2020 0 repositories listed
-
A Deep Reinforcement Learning Framework for Continuous Intraday Market Bidding13 Apr 2020 0 repositories listed
-
A non-cooperative meta-modeling game for automated third-party calibrating, validating, and falsifying constitutive laws with parallelized adversarial attacks13 Apr 2020 0 repositories listed
-
Aspect and Opinion Aware Abstractive Review Summarization with Reinforced Hard Typed Decoder13 Apr 2020 0 repositories listed
-
K-spin Hamiltonian for quantum-resolvable Markov decision processes13 Apr 2020 0 repositories listed
-
Reinforced Curriculum Learning on Pre-trained Neural Machine Translation Models13 Apr 2020 0 repositories listed
-
Thinking While Moving: Deep Reinforcement Learning with Concurrent Control13 Apr 2020 0 repositories listed
-
Reinforcement Learning via Reasoning from Demonstration12 Apr 2020 0 repositories listed
-
Certifiable Robustness to Adversarial State Uncertainty in Deep Reinforcement Learning11 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning for Process Control: A Primer for Beginners11 Apr 2020 0 repositories listed
-
Reinforcement Learning via Gaussian Processes with Neural Network Dual Kernels10 Apr 2020 0 repositories listed
-
Deep Reinforcement Learning (DRL): Another Perspective for Unsupervised Wireless Localization9 Apr 2020 0 repositories listed
-
Learning to Drive Off Road on Smooth Terrain in Unstructured Environments Using an On-Board Camera and Sparse Aerial Images9 Apr 2020 0 repositories listed
-
On Linear Stochastic Approximation: Fine-grained Polyak-Ruppert and Non-Asymptotic Concentration9 Apr 2020 0 repositories listed
-
Policy Gradient using Weak Derivatives for Reinforcement Learning9 Apr 2020 0 repositories listed
-
Quantifying the Impact of Non-Stationarity in Reinforcement Learning-Based Traffic Signal Control9 Apr 2020 0 repositories listed
-
Re-conceptualising the Language Game Paradigm in the Framework of Multi-Agent Reinforcement Learning9 Apr 2020 0 repositories listed
-
Reinforced Anytime Bottom Up Rule Learning for Knowledge Graph Completion9 Apr 2020 0 repositories listed
-
Risk-Aware High-level Decisions for Automated Driving at Occluded Intersections with Reinforcement Learning9 Apr 2020 0 repositories listed
-
Adaptive Stress Testing without Domain Heuristics using Go-Explore8 Apr 2020 0 repositories listed
-
GeneCAI: Genetic Evolution for Acquiring Compact AI8 Apr 2020 0 repositories listed
-
Monte-Carlo Siamese Policy on Actor for Satellite Image Super Resolution8 Apr 2020 0 repositories listed
-
Resource Management for Blockchain-enabled Federated Learning: A Deep Reinforcement Learning Approach8 Apr 2020 0 repositories listed
-
How Do You Act? An Empirical Study to Understand Behavior of Deep Reinforcement Learning Agents7 Apr 2020 0 repositories listed
-
Online Constrained Model-based Reinforcement Learning7 Apr 2020 0 repositories listed
-
Practical Data Poisoning Attack against Next-Item Recommendation7 Apr 2020 0 repositories listed
-
Optimistic Agent: Accurate Graph-Based Value Estimation for More Successful Visual Navigation7 Apr 2020 0 repositories listed
-
B-SCST: Bayesian Self-Critical Sequence Training for Image Captioning6 Apr 2020 0 repositories listed
-
CNN2Gate: Toward Designing a General Framework for Implementation of Convolutional Neural Networks on FPGA6 Apr 2020 0 repositories listed
-
Intrinsic Exploration as Multi-Objective RL6 Apr 2020 0 repositories listed
-
Networked Multi-Agent Reinforcement Learning with Emergent Communication6 Apr 2020 0 repositories listed
-
Technical Report: Adaptive Control for Linearizable Systems Using On-Policy Reinforcement Learning6 Apr 2020 0 repositories listed
-
Uniform State Abstraction For Reinforcement Learning6 Apr 2020 0 repositories listed
-
Using Generative Adversarial Nets on Atari Games for Feature Extraction in Deep Reinforcement Learning6 Apr 2020 0 repositories listed
-
Weakly-Supervised Reinforcement Learning for Controllable Behavior6 Apr 2020 0 repositories listed
-
Zero-Shot Learning of Text Adventure Games with Sentence-Level Semantics6 Apr 2020 0 repositories listed
-
Multi-agent Reinforcement Learning for Resource Allocation in IoT networks with Edge Computing5 Apr 2020 0 repositories listed
-
Reinforced Multi-task Approach for Multi-hop Question Generation5 Apr 2020 0 repositories listed
-
Reinforcement Learning Architectures: SAC, TAC, and ESAC5 Apr 2020 0 repositories listed
-
Stylistic Dialogue Generation via Information-Guided Reinforcement Learning Strategy5 Apr 2020 0 repositories listed
-
A Deep Ensemble Multi-Agent Reinforcement Learning Approach for Air Traffic Control3 Apr 2020 0 repositories listed
-
Reinforcement Learning for Mixed-Integer Problems Based on MPC3 Apr 2020 0 repositories listed
-
Average Reward Adjusted Discounted Reinforcement Learning: Near-Blackwell-Optimal Policies for Real-World Applications2 Apr 2020 0 repositories listed
-
Continuous Motion Planning with Temporal Logic Specifications using Deep Neural Networks2 Apr 2020 0 repositories listed
-
Exploration of Reinforcement Learning for Event Camera using Car-like Robots2 Apr 2020 0 repositories listed
-
Learning Agile Robotic Locomotion Skills by Imitating Animals2 Apr 2020 0 repositories listed
-
Automated Quantification of CT Patterns Associated with COVID-19 from Chest CT2 Apr 2020 0 repositories listed
-
Safe Reinforcement Learning via Projection on a Safe Set: How to Achieve Optimality?2 Apr 2020 0 repositories listed
-
Value Driven Representation for Human-in-the-Loop Reinforcement Learning2 Apr 2020 0 repositories listed
-
A New Challenge: Approaching Tetris Link with AI1 Apr 2020 0 repositories listed
-
Constrained-Space Optimization and Reinforcement Learning for Complex Tasks1 Apr 2020 0 repositories listed
-
Counterfactual Multi-Agent Reinforcement Learning with Graph Convolution Communication1 Apr 2020 0 repositories listed
-
Development of swarm behavior in artificial learning agents that adapt to different foraging environments1 Apr 2020 0 repositories listed
-
Statistically Model Checking PCTL Specifications on Markov Decision Processes via Reinforcement Learning1 Apr 2020 0 repositories listed
-
Work in Progress: Temporally Extended Auxiliary Tasks1 Apr 2020 0 repositories listed
-
Controlling Rayleigh-Bénard convection via Reinforcement Learning31 Mar 2020 0 repositories listed
-
Enhanced Rolling Horizon Evolution Algorithm with Opponent Model Learning: Results for the Fighting Game AI Competition31 Mar 2020 0 repositories listed
-
Leverage the Average: an Analysis of KL Regularization in RL31 Mar 2020 0 repositories listed
-
Mimicking Evolution with Reinforcement Learning31 Mar 2020 0 repositories listed
-
Optimal Bidding Strategy without Exploration in Real-time Bidding31 Mar 2020 0 repositories listed
-
Robotic Table Tennis with Model-Free Reinforcement Learning31 Mar 2020 0 repositories listed
-
Continual Learning with Node-Importance based Adaptive Group Sparse Regularization30 Mar 2020 0 repositories listed
-
Model-Reference Reinforcement Learning Control of Autonomous Surface Vehicles with Uncertainties30 Mar 2020 0 repositories listed
-
Stochastic Flows and Geometric Optimization on the Orthogonal Group30 Mar 2020 0 repositories listed
-
Learning and Testing Variable Partitions29 Mar 2020 0 repositories listed
-
Parallel Knowledge Transfer in Multi-Agent Reinforcement Learning29 Mar 2020 0 repositories listed
-
When Autonomous Systems Meet Accuracy and Transferability through AI: A Survey29 Mar 2020 0 repositories listed
-
Learning medical triage from clinicians using Deep Q-Learning28 Mar 2020 0 repositories listed
-
Streamlined Empirical Bayes Fitting of Linear Mixed Models in Mobile Health28 Mar 2020 0 repositories listed
-
A Distributional Analysis of Sampling-Based Reinforcement Learning Algorithms27 Mar 2020 0 repositories listed
-
Adaptive Reward-Poisoning Attacks against Reinforcement Learning27 Mar 2020 0 repositories listed
-
AirRL: A Reinforcement Learning Approach to Urban Air Quality Inference27 Mar 2020 0 repositories listed
-
Towards Better Opioid Antagonists Using Deep Reinforcement Learning26 Mar 2020 0 repositories listed
-
ACNMP: Skill Transfer and Task Extrapolation through Learning from Demonstration and Reinforcement Learning via Representation Sharing25 Mar 2020 0 repositories listed
-
AliExpress Learning-To-Rank: Maximizing Online Model Performance without Going Online25 Mar 2020 0 repositories listed
-
Convergence of Recursive Stochastic Algorithms using Wasserstein Divergence25 Mar 2020 0 repositories listed
-
Black-box Off-policy Estimation for Infinite-Horizon Reinforcement Learning24 Mar 2020 0 repositories listed
-
Distributional Reinforcement Learning with Ensembles24 Mar 2020 0 repositories listed
-
Driver Modeling through Deep Reinforcement Learning and Behavioral Game Theory24 Mar 2020 0 repositories listed
-
Finite-Time Analysis of Stochastic Gradient Descent under Markov Randomness24 Mar 2020 0 repositories listed
-
Learn to Schedule (LEASCH): A Deep reinforcement learning approach for radio resource scheduling in the 5G MAC layer24 Mar 2020 0 repositories listed
-
Learning Compact Reward for Image Captioning24 Mar 2020 0 repositories listed
-
Learning to Play Soccer by Reinforcement and Applying Sim-to-Real to Compete in the Real World24 Mar 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.