Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 120
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 120 of 152: papers 11,901 to 12,000 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
The Effect of Multi-step Methods on Overestimation in Deep Reinforcement Learning23 Jun 2020 0 repositories listed
-
Accelerated Deep Reinforcement Learning Based Load Shedding for Emergency Voltage Control22 Jun 2020 0 repositories listed
-
Constrained Combinatorial Optimization with Reinforcement Learning22 Jun 2020 0 repositories listed
-
Ecological Reinforcement Learning22 Jun 2020 0 repositories listed
-
Efficient Sampling-Based Maximum Entropy Inverse Reinforcement Learning with Application to Autonomous Driving22 Jun 2020 0 repositories listed
-
Near-Optimal Reinforcement Learning with Self-Play22 Jun 2020 0 repositories listed
-
Provably Efficient Causal Reinforcement Learning with Confounded Observational Data22 Jun 2020 0 repositories listed
-
QTRAN++: Improved Value Transformation for Cooperative Multi-Agent Reinforcement Learning22 Jun 2020 0 repositories listed
-
Risk-Sensitive Reinforcement Learning: Near-Optimal Risk-Sample Tradeoff in Regret22 Jun 2020 0 repositories listed
-
Sample-Efficient Reinforcement Learning of Undercomplete POMDPs22 Jun 2020 0 repositories listed
-
Breaking the Curse of Many Agents: Provable Mean Embedding Q-Iteration for Mean-Field Reinforcement Learning21 Jun 2020 0 repositories listed
-
Gradient-EM Bayesian Meta-learning21 Jun 2020 0 repositories listed
-
Hierarchical Reinforcement Learning for Deep Goal Reasoning: An Expressiveness Analysis21 Jun 2020 0 repositories listed
-
Reinforcement Learning for Mean Field Games with Strategic Complementarities21 Jun 2020 0 repositories listed
-
Off-Policy Self-Critical Training for Transformer in Visual Paragraph Generation21 Jun 2020 0 repositories listed
-
Towards Tractable Optimism in Model-Based Reinforcement Learning21 Jun 2020 0 repositories listed
-
Accelerating Safe Reinforcement Learning with Constraint-mismatched Policies20 Jun 2020 0 repositories listed
-
Entropic Risk Constrained Soft-Robust Policy Optimization20 Jun 2020 0 repositories listed
-
Langevin Dynamics for Adaptive Inverse Reinforcement Learning of Stochastic Gradient Algorithms20 Jun 2020 0 repositories listed
-
Robust Reinforcement Learning using Least Squares Policy Iteration with Provable Performance Guarantees20 Jun 2020 0 repositories listed
-
A Reinforcement Learning Approach for Transient Control of Liquid Rocket Engines19 Jun 2020 0 repositories listed
-
Learn to Earn: Enabling Coordination within a Ride Hailing Fleet19 Jun 2020 0 repositories listed
-
NROWAN-DQN: A Stable Noisy Network with Noise Reduction and Online Weight Adjustment for Exploration19 Jun 2020 0 repositories listed
-
On Reward-Free Reinforcement Learning with Linear Function Approximation19 Jun 2020 0 repositories listed
-
FISAR: Forward Invariant Safe Reinforcement Learning with a Deep Neural Network-Based Optimize19 Jun 2020 0 repositories listed
-
Cooperative Multi-Agent Reinforcement Learning with Partial Observations18 Jun 2020 0 repositories listed
-
Deep Reinforcement Learning amidst Lifelong Non-Stationarity18 Jun 2020 0 repositories listed
-
Distributed Value Function Approximation for Collaborative Multi-Agent Reinforcement Learning18 Jun 2020 0 repositories listed
-
FLAMBE: Structural Complexity and Representation Learning of Low Rank MDPs18 Jun 2020 0 repositories listed
-
Interactive Recommender System via Knowledge Graph-enhanced Reinforcement Learning18 Jun 2020 0 repositories listed
-
Provably adaptive reinforcement learning in metric spaces18 Jun 2020 0 repositories listed
-
WD3: Taming the Estimation Bias in Deep Reinforcement Learning18 Jun 2020 0 repositories listed
-
Deep Reinforcement Learning Controller for 3D Path-following and Collision Avoidance by Autonomous Underwater Vehicles17 Jun 2020 0 repositories listed
-
Eco-Vehicular Edge Networks for Connected Transportation: A Distributed Multi-Agent Reinforcement Learning Approach17 Jun 2020 0 repositories listed
-
Introduction to Machine Learning for Accelerator Physics17 Jun 2020 0 repositories listed
-
Parameterized MDPs and Reinforcement Learning Problems -- A Maximum Entropy Principle Based Framework17 Jun 2020 0 repositories listed
-
Policy Evaluation and Seeking for Multi-Agent Reinforcement Learning via Best Response17 Jun 2020 0 repositories listed
-
Reinforcement Learning with Uncertainty Estimation for Tactical Decision-Making in Intersections17 Jun 2020 0 repositories listed
-
COLREG-Compliant Collision Avoidance for Unmanned Surface Vehicle using Deep Reinforcement Learning16 Jun 2020 0 repositories listed
-
Index Selection for NoSQL Database with Deep Reinforcement Learning16 Jun 2020 0 repositories listed
-
Model Embedding Model-Based Reinforcement Learning16 Jun 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning for Adaptive User Association in Dynamic mmWave Networks16 Jun 2020 0 repositories listed
-
Online Reinforcement Learning Control by Direct Heuristic Dynamic Programming: from Time-Driven to Event-Driven16 Jun 2020 0 repositories listed
-
Preference-based Reinforcement Learning with Finite-Time Guarantees16 Jun 2020 0 repositories listed
-
Reinforcement Learning Control of Robotic Knee with Human in the Loop by Flexible Policy Iteration16 Jun 2020 0 repositories listed
-
RL-CycleGAN: Reinforcement Learning Aware Simulation-To-Real16 Jun 2020 0 repositories listed
-
ShieldNN: A Provably Safe NN Filter for Unsafe NN Controllers16 Jun 2020 0 repositories listed
-
Solving the Order Batching and Sequencing Problem using Deep Reinforcement Learning16 Jun 2020 0 repositories listed
-
Task-agnostic Exploration in Reinforcement Learning16 Jun 2020 0 repositories listed
-
The Sample Complexity of Teaching-by-Reinforcement on Q-Learning16 Jun 2020 0 repositories listed
-
An online evolving framework for advancing reinforcement-learning based automated vehicle control15 Jun 2020 0 repositories listed
-
Designing high-fidelity multi-qubit gates for semiconductor quantum dots through deep reinforcement learning15 Jun 2020 0 repositories listed
-
Runtime Adaptation in Wireless Sensor Nodes Using Structured Learning15 Jun 2020 0 repositories listed
-
Variable Gain Gradient Descent-based Reinforcement Learning for Robust Optimal Tracking Control of Uncertain Nonlinear System with Input-Constraints15 Jun 2020 0 repositories listed
-
Adversarial Attacks and Detection on Reinforcement Learning-Based Interactive Recommender Systems14 Jun 2020 0 repositories listed
-
Non-local Policy Optimization via Diversity-regularized Collaborative Exploration14 Jun 2020 0 repositories listed
-
Reinforcement Learning with Supervision from Noisy Demonstrations14 Jun 2020 0 repositories listed
-
Tackling Morpion Solitaire with AlphaZero-likeRanked Reward Reinforcement Learning14 Jun 2020 0 repositories listed
-
Hindsight Expectation Maximization for Goal-conditioned Reinforcement Learning13 Jun 2020 0 repositories listed
-
Reinforcement Learning as Iterative and Amortised Inference13 Jun 2020 0 repositories listed
-
A Brief Look at Generalization in Visual Meta-Reinforcement Learning12 Jun 2020 0 repositories listed
-
Bridging Worlds in Reinforcement Learning with Model-Advantage12 Jun 2020 0 repositories listed
-
Continuous Control for Searching and Planning with a Learned Model12 Jun 2020 0 repositories listed
-
Decorrelated Double Q-learning12 Jun 2020 0 repositories listed
-
Deep Reinforcement Learning for Neural Control12 Jun 2020 0 repositories listed
-
Explore then Execute: Adapting without Rewards via Factorized Meta-Reinforcement Learning12 Jun 2020 0 repositories listed
-
Generalizing Curricula for Reinforcement Learning12 Jun 2020 0 repositories listed
-
Hierarchical reinforcement learning for efficent exploration and transfer12 Jun 2020 0 repositories listed
-
Human and Multi-Agent collaboration in a human-MARL teaming framework12 Jun 2020 0 repositories listed
-
Learning Intrinsically Motivated Options to Stimulate Policy Exploration12 Jun 2020 0 repositories listed
-
Logical Composition in Lifelong Reinforcement Learning12 Jun 2020 0 repositories listed
-
Meta-Reinforcement Learning Robust to Distributional Shift via Model Identification and Experience Relabeling12 Jun 2020 0 repositories listed
-
Potential Field Guided Actor-Critic Reinforcement Learning12 Jun 2020 0 repositories listed
-
Safety-guaranteed Reinforcement Learning based on Multi-class Support Vector Machine12 Jun 2020 0 repositories listed
-
StarCraft II Build Order Optimization using Deep Reinforcement Learning and Monte-Carlo Tree Search12 Jun 2020 0 repositories listed
-
Systematic Generalisation through Task Temporal Logic and Deep Reinforcement Learning12 Jun 2020 0 repositories listed
-
Using Reinforcement Learning to Allocate and Manage Service Function Chains in Cellular Networks12 Jun 2020 0 repositories listed
-
Deep Reinforcement Learning for Electric Transmission Voltage Control11 Jun 2020 0 repositories listed
-
Exploration by Maximizing Rényi Entropy for Reward-Free RL Framework11 Jun 2020 0 repositories listed
-
Multi-Agent Informational Learning Processes11 Jun 2020 0 repositories listed
-
Sample Efficient Reinforcement Learning via Low-Rank Matrix Estimation11 Jun 2020 0 repositories listed
-
Scalable Multi-Agent Reinforcement Learning for Networked Systems with Average Reward11 Jun 2020 0 repositories listed
-
Surveys without Questions: A Reinforcement Learning Approach11 Jun 2020 0 repositories listed
-
Zeroth-Order Supervised Policy Improvement11 Jun 2020 0 repositories listed
-
Deep reinforcement learning for optical systems: A case study of mode-locked lasers10 Jun 2020 0 repositories listed
-
Development of A Stochastic Traffic Environment with Generative Time-Series Models for Improving Generalization Capabilities of Autonomous Driving Agents10 Jun 2020 0 repositories listed
-
Learning to Play Table Tennis From Scratch using Muscular Robots10 Jun 2020 0 repositories listed
-
Machine learning and control engineering: The model-free case10 Jun 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning in a Realistic Limit Order Book Market Simulation10 Jun 2020 0 repositories listed
-
Off-Policy Risk-Sensitive Reinforcement Learning Based Constrained Robust Optimal Control10 Jun 2020 0 repositories listed
-
Privacy-Cost Management in Smart Meters with Mutual Information-Based Reinforcement Learning10 Jun 2020 0 repositories listed
-
Q-greedyUCB: a New Exploration Policy for Adaptive and Resource-efficient Scheduling10 Jun 2020 0 repositories listed
-
Searching Learning Strategy with Reinforcement Learning for 3D Medical Image Segmentation10 Jun 2020 0 repositories listed
-
Self-Supervised Reinforcement Learning for Recommender Systems10 Jun 2020 0 repositories listed
-
Transient Non-Stationarity and Generalisation in Deep Reinforcement Learning10 Jun 2020 0 repositories listed
-
An overall view of key problems in algorithmic trading and recent progress9 Jun 2020 0 repositories listed
-
Causal Discovery from Incomplete Data using An Encoder and Reinforcement Learning9 Jun 2020 0 repositories listed
-
Distributed Learning on Heterogeneous Resource-Constrained Devices9 Jun 2020 0 repositories listed
-
Policy-focused Agent-based Modeling using RL Behavioral Models9 Jun 2020 0 repositories listed
-
Stealing Deep Reinforcement Learning Models for Fun and Profit9 Jun 2020 0 repositories listed