Browse State-of-the-Art › Reinforcement Learning › Papers, page 70
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 70 of 132: papers 6,901 to 7,000 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Redefining Counterfactual Explanations for Reinforcement Learning: Overview, Challenges and Opportunities21 Oct 2022 0 repositories listed
-
Deep Reinforcement Learning for Inverse Inorganic Materials Design21 Oct 2022 0 repositories listed
-
Group Distributionally Robust Reinforcement Learning with Hierarchical Latent Variables21 Oct 2022 0 repositories listed
-
Implicit Offline Reinforcement Learning via Supervised Learning21 Oct 2022 0 repositories listed
-
Integrating Policy Summaries with Reward Decomposition for Explaining Reinforcement Learning Agents21 Oct 2022 0 repositories listed
-
On the connection between Bregman divergence and value in regularized Markov decision processes21 Oct 2022 0 repositories listed
-
Robust Imitation via Mirror Descent Inverse Reinforcement Learning20 Oct 2022 0 repositories listed
-
Hierarchical Reinforcement Learning for Furniture Layout in Virtual Indoor Scenes19 Oct 2022 0 repositories listed
-
Oracles & Followers: Stackelberg Equilibria in Deep Multi-Agent Reinforcement Learning19 Oct 2022 0 repositories listed
-
Palm up: Playing in the Latent Manifold for Unsupervised Pretraining19 Oct 2022 0 repositories listed
-
Provably Safe Reinforcement Learning via Action Projection using Reachability Analysis and Polynomial Zonotopes19 Oct 2022 0 repositories listed
-
Scaling Laws for Reward Model Overoptimization19 Oct 2022 0 repositories listed
-
RPM: Generalizable Behaviors for Multi-Agent Reinforcement Learning18 Oct 2022 0 repositories listed
-
Unpacking Reward Shaping: Understanding the Benefits of Reward Engineering on Sample Complexity18 Oct 2022 0 repositories listed
-
Boosting Offline Reinforcement Learning via Data Rebalancing17 Oct 2022 0 repositories listed
-
You Only Live Once: Single-Life Reinforcement Learning17 Oct 2022 0 repositories listed
-
Data-Efficient Pipeline for Offline Reinforcement Learning with Limited Data16 Oct 2022 0 repositories listed
-
Entropy Regularized Reinforcement Learning with Cascading Networks16 Oct 2022 0 repositories listed
-
The Impact of Task Underspecification in Evaluating Deep Reinforcement Learning16 Oct 2022 0 repositories listed
-
Towards an Interpretable Hierarchical Agent Framework using Semantic Goals16 Oct 2022 0 repositories listed
-
DyFEn: Agent-Based Fee Setting in Payment Channel Networks15 Oct 2022 0 repositories listed
-
Near-Optimal Regret Bounds for Multi-batch Reinforcement Learning15 Oct 2022 0 repositories listed
-
Reinforcement Learning for ConnectX15 Oct 2022 0 repositories listed
-
A Reinforcement Learning Approach to Estimating Long-term Treatment Effects14 Oct 2022 0 repositories listed
-
A Scalable Finite Difference Method for Deep Reinforcement Learning14 Oct 2022 0 repositories listed
-
Adaptive patch foraging in deep reinforcement learning agents14 Oct 2022 0 repositories listed
-
Multi-trainer Interactive Reinforcement Learning System14 Oct 2022 0 repositories listed
-
Robust Preference Learning for Storytelling via Contrastive Reinforcement Learning14 Oct 2022 0 repositories listed
-
A Concise Introduction to Reinforcement Learning in Robotics13 Oct 2022 0 repositories listed
-
Causality-driven Hierarchical Structure Discovery for Reinforcement Learning13 Oct 2022 0 repositories listed
-
Efficient circuit implementation for coined quantum walks on binary trees and application to reinforcement learning13 Oct 2022 0 repositories listed
-
Object-Category Aware Reinforcement Learning13 Oct 2022 0 repositories listed
-
Observed Adversaries in Deep Reinforcement Learning13 Oct 2022 0 repositories listed
-
Optimal Control of Material Micro-Structures13 Oct 2022 0 repositories listed
-
Personalized Federated Hypernetworks for Privacy Preservation in Multi-Task Reinforcement Learning13 Oct 2022 0 repositories listed
-
Reinforcement Learning with Unbiased Policy Evaluation and Linear Function Approximation13 Oct 2022 0 repositories listed
-
DQLAP: Deep Q-Learning Recommender Algorithm with Update Policy for a Real Steam Turbine System12 Oct 2022 0 repositories listed
-
Explaining Online Reinforcement Learning Decisions of Self-Adaptive Systems12 Oct 2022 0 repositories listed
-
Real World Offline Reinforcement Learning with Realistic Data Source12 Oct 2022 0 repositories listed
-
Reinforcement Learning with Automated Auxiliary Loss Search12 Oct 2022 0 repositories listed
-
Smooth Trajectory Collision Avoidance through Deep Reinforcement Learning12 Oct 2022 0 repositories listed
-
Broad-persistent Advice for Interactive Reinforcement Learning Scenarios11 Oct 2022 0 repositories listed
-
Multi-User Reinforcement Learning with Low Rank Rewards11 Oct 2022 0 repositories listed
-
Regret Bounds for Risk-Sensitive Reinforcement Learning11 Oct 2022 0 repositories listed
-
The Role of Exploration for Task Transfer in Reinforcement Learning11 Oct 2022 0 repositories listed
-
Creating a Dynamic Quadrupedal Robotic Goalkeeper with Reinforcement Learning10 Oct 2022 0 repositories listed
-
Long N-step Surrogate Stage Reward to Reduce Variances of Deep Reinforcement Learning in Complex Problems10 Oct 2022 0 repositories listed
-
Towards a Theoretical Foundation of Policy Optimization for Learning Control Policies10 Oct 2022 0 repositories listed
-
The Role of Coverage in Online Reinforcement Learning9 Oct 2022 0 repositories listed
-
Advice Conformance Verification by Reinforcement Learning agents for Human-in-the-Loop7 Oct 2022 0 repositories listed
-
Algorithmic Trading Using Continuous Action Space Deep Reinforcement Learning7 Oct 2022 0 repositories listed
-
Reinforcement Learning Approach for Multi-Agent Flexible Scheduling Problems7 Oct 2022 0 repositories listed
-
Deep Inventory Management6 Oct 2022 0 repositories listed
-
Digital Human Interactive Recommendation Decision-Making Based on Reinforcement Learning6 Oct 2022 0 repositories listed
-
6 Oct 2022 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Learning Algorithms for Intelligent Agents and Mechanisms6 Oct 2022 0 repositories listed
-
Low-Thrust Orbital Transfer using Dynamics-Agnostic Reinforcement Learning6 Oct 2022 0 repositories listed
-
Meta Reinforcement Learning for Optimal Design of Legged Robots6 Oct 2022 0 repositories listed
-
Reinforcement Learning with Large Action Spaces for Neural Machine Translation6 Oct 2022 0 repositories listed
-
Neural Distillation as a State Representation Bottleneck in Reinforcement Learning5 Oct 2022 0 repositories listed
-
On Neural Consolidation for Transfer in Reinforcement Learning5 Oct 2022 0 repositories listed
-
Using Deep Reinforcement Learning for mmWave Real-Time Scheduling4 Oct 2022 0 repositories listed
-
Handling Sparse Rewards in Reinforcement Learning Using Model Predictive Control4 Oct 2022 0 repositories listed
-
Hyperbolic Deep Reinforcement Learning4 Oct 2022 0 repositories listed
-
Learning Dynamic Abstract Representations for Sample-Efficient Reinforcement Learning4 Oct 2022 0 repositories listed
-
Maximum-Likelihood Inverse Reinforcement Learning with Finite-Time Guarantees4 Oct 2022 0 repositories listed
-
CostNet: An End-to-End Framework for Goal-Directed Reinforcement Learning3 Oct 2022 0 repositories listed
-
Interpretable Option Discovery using Deep Q-Learning and Variational Autoencoders3 Oct 2022 0 repositories listed
-
MSRL: Distributed Reinforcement Learning with Dataflow Fragments3 Oct 2022 0 repositories listed
-
Offline Reinforcement Learning with Differentiable Function Approximation is Provably Efficient3 Oct 2022 0 repositories listed
-
Policy Gradient for Reinforcement Learning with General Utilities3 Oct 2022 0 repositories listed
-
Policy Gradients for Probabilistic Constrained Reinforcement Learning2 Oct 2022 0 repositories listed
-
Robust Bayesian optimization with reinforcement learned acquisition functions2 Oct 2022 0 repositories listed
-
Comparing BERT-based Reward Functions for Deep Reinforcement Learning in Machine Translation1 Oct 2022 0 repositories listed
-
Parsing Natural Language into Propositional and First-Order Logic with Dual Reinforcement Learning1 Oct 2022 0 repositories listed
-
ASPiRe:Adaptive Skill Priors for Reinforcement Learning30 Sep 2022 0 repositories listed
-
Efficiently Learning Small Policies for Locomotion and Manipulation30 Sep 2022 0 repositories listed
-
Bounded Robustness in Reinforcement Learning via Lexicographic Objectives30 Sep 2022 0 repositories listed
-
Programmable Control of Ultrasound Swarmbots through Reinforcement Learning30 Sep 2022 0 repositories listed
-
Blessing from Human-AI Interaction: Super Reinforcement Learning in Confounded Environments29 Sep 2022 0 repositories listed
-
Contrastive Unsupervised Learning of World Model with Invariant Causal Features29 Sep 2022 0 repositories listed
-
Ensemble Reinforcement Learning in Continuous Spaces -- A Hierarchical Multi-Step Approach for Policy Training29 Sep 2022 0 repositories listed
-
Learning Parsimonious Dynamics for Generalization in Reinforcement Learning29 Sep 2022 0 repositories listed
-
Online Weighted Q-Ensembles for Reduced Hyperparameter Tuning in Reinforcement Learning29 Sep 2022 0 repositories listed
-
Reinforcement Learning Algorithms: An Overview and Classification29 Sep 2022 0 repositories listed
-
Argumentative Reward Learning: Reasoning About Human Preferences28 Sep 2022 0 repositories listed
-
Disentangling Transfer in Continual Reinforcement Learning28 Sep 2022 0 repositories listed
-
Predictive Crypto-Asset Automated Market Making Architecture for Decentralized Finance using Deep Reinforcement Learning28 Sep 2022 0 repositories listed
-
DCE: Offline Reinforcement Learning With Double Conservative Estimates27 Sep 2022 0 repositories listed
-
Reinforcement Learning with Non-Exponential Discounting27 Sep 2022 0 repositories listed
-
Safe Reinforcement Learning of Dynamic High-Dimensional Robotic Tasks: Navigation, Manipulation, Interaction27 Sep 2022 0 repositories listed
-
DEFT: Diverse Ensembles for Fast Transfer in Reinforcement Learning26 Sep 2022 0 repositories listed
-
Delayed Geometric Discounts: An Alternative Criterion for Reinforcement Learning26 Sep 2022 0 repositories listed
-
Overcoming Referential Ambiguity in Language-Guided Goal-Conditioned Reinforcement Learning26 Sep 2022 0 repositories listed
-
Deep Reinforcement Learning for Adaptive Mesh Refinement25 Sep 2022 0 repositories listed
-
Reward Learning using Structural Motifs in Inverse Reinforcement Learning25 Sep 2022 0 repositories listed
-
Fast Lifelong Adaptive Inverse Reinforcement Learning from Demonstrations24 Sep 2022 0 repositories listed
-
Quantification before Selection: Active Dynamics Preference for Robust Reinforcement Learning23 Sep 2022 0 repositories listed
-
Minimizing Human Assistance: Augmenting a Single Demonstration for Deep Reinforcement Learning22 Sep 2022 0 repositories listed
-
Parallel Reinforcement Learning Simulation for Visual Quadrotor Navigation22 Sep 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.