Browse State-of-the-Art › reinforcement-learning › Papers, page 75
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 75 of 135: papers 7,401 to 7,500 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
The Impact of Task Underspecification in Evaluating Deep Reinforcement Learning16 Oct 2022 0 repositories listed
-
Towards an Interpretable Hierarchical Agent Framework using Semantic Goals16 Oct 2022 0 repositories listed
-
DyFEn: Agent-Based Fee Setting in Payment Channel Networks15 Oct 2022 0 repositories listed
-
Near-Optimal Regret Bounds for Multi-batch Reinforcement Learning15 Oct 2022 0 repositories listed
-
PI-QT-Opt: Predictive Information Improves Multi-Task Robotic Reinforcement Learning at Scale15 Oct 2022 0 repositories listed
-
Reinforcement Learning for ConnectX15 Oct 2022 0 repositories listed
-
A Reinforcement Learning Approach to Estimating Long-term Treatment Effects14 Oct 2022 0 repositories listed
-
A Scalable Finite Difference Method for Deep Reinforcement Learning14 Oct 2022 0 repositories listed
-
Query Rewriting for Effective Misinformation Discovery14 Oct 2022 0 repositories listed
-
Adaptive patch foraging in deep reinforcement learning agents14 Oct 2022 0 repositories listed
-
Multi-trainer Interactive Reinforcement Learning System14 Oct 2022 0 repositories listed
-
Robust Preference Learning for Storytelling via Contrastive Reinforcement Learning14 Oct 2022 0 repositories listed
-
A Concise Introduction to Reinforcement Learning in Robotics13 Oct 2022 0 repositories listed
-
Causality-driven Hierarchical Structure Discovery for Reinforcement Learning13 Oct 2022 0 repositories listed
-
Efficient circuit implementation for coined quantum walks on binary trees and application to reinforcement learning13 Oct 2022 0 repositories listed
-
Object-Category Aware Reinforcement Learning13 Oct 2022 0 repositories listed
-
Observed Adversaries in Deep Reinforcement Learning13 Oct 2022 0 repositories listed
-
Optimal Control of Material Micro-Structures13 Oct 2022 0 repositories listed
-
Output Feedback Adaptive Optimal Control of Affine Nonlinear systems with a Linear Measurement Model13 Oct 2022 0 repositories listed
-
Personalized Federated Hypernetworks for Privacy Preservation in Multi-Task Reinforcement Learning13 Oct 2022 0 repositories listed
-
Reinforcement Learning with Unbiased Policy Evaluation and Linear Function Approximation13 Oct 2022 0 repositories listed
-
Towards Multi-Agent Reinforcement Learning driven Over-The-Counter Market Simulations13 Oct 2022 0 repositories listed
-
DQLAP: Deep Q-Learning Recommender Algorithm with Update Policy for a Real Steam Turbine System12 Oct 2022 0 repositories listed
-
Explaining Online Reinforcement Learning Decisions of Self-Adaptive Systems12 Oct 2022 0 repositories listed
-
Real World Offline Reinforcement Learning with Realistic Data Source12 Oct 2022 0 repositories listed
-
Reinforcement Learning with Automated Auxiliary Loss Search12 Oct 2022 0 repositories listed
-
Smooth Trajectory Collision Avoidance through Deep Reinforcement Learning12 Oct 2022 0 repositories listed
-
Broad-persistent Advice for Interactive Reinforcement Learning Scenarios11 Oct 2022 0 repositories listed
-
Edge-Cloud Cooperation for DNN Inference via Reinforcement Learning and Supervised Learning11 Oct 2022 0 repositories listed
-
Multi-User Reinforcement Learning with Low Rank Rewards11 Oct 2022 0 repositories listed
-
Regret Bounds for Risk-Sensitive Reinforcement Learning11 Oct 2022 0 repositories listed
-
The Role of Exploration for Task Transfer in Reinforcement Learning11 Oct 2022 0 repositories listed
-
Creating a Dynamic Quadrupedal Robotic Goalkeeper with Reinforcement Learning10 Oct 2022 0 repositories listed
-
Long N-step Surrogate Stage Reward to Reduce Variances of Deep Reinforcement Learning in Complex Problems10 Oct 2022 0 repositories listed
-
Simulating Coverage Path Planning with Roomba10 Oct 2022 0 repositories listed
-
Towards a Theoretical Foundation of Policy Optimization for Learning Control Policies10 Oct 2022 0 repositories listed
-
Equivalence of Optimality Criteria for Markov Decision Process and Model Predictive Control9 Oct 2022 0 repositories listed
-
The Role of Coverage in Online Reinforcement Learning9 Oct 2022 0 repositories listed
-
Advice Conformance Verification by Reinforcement Learning agents for Human-in-the-Loop7 Oct 2022 0 repositories listed
-
Algorithmic Trading Using Continuous Action Space Deep Reinforcement Learning7 Oct 2022 0 repositories listed
-
Multi-agent Deep Covering Skill Discovery7 Oct 2022 0 repositories listed
-
Reinforcement Learning Approach for Multi-Agent Flexible Scheduling Problems7 Oct 2022 0 repositories listed
-
Deep Inventory Management6 Oct 2022 0 repositories listed
-
Digital Human Interactive Recommendation Decision-Making Based on Reinforcement Learning6 Oct 2022 0 repositories listed
-
6 Oct 2022 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Learning Algorithms for Intelligent Agents and Mechanisms6 Oct 2022 0 repositories listed
-
Low-Thrust Orbital Transfer using Dynamics-Agnostic Reinforcement Learning6 Oct 2022 0 repositories listed
-
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning6 Oct 2022 0 repositories listed
-
Meta Reinforcement Learning for Optimal Design of Legged Robots6 Oct 2022 0 repositories listed
-
Reinforcement Learning with Large Action Spaces for Neural Machine Translation6 Oct 2022 0 repositories listed
-
A Novel Entropy-Maximizing TD3-based Reinforcement Learning for Automatic PID Tuning5 Oct 2022 0 repositories listed
-
Neural Distillation as a State Representation Bottleneck in Reinforcement Learning5 Oct 2022 0 repositories listed
-
On Neural Consolidation for Transfer in Reinforcement Learning5 Oct 2022 0 repositories listed
-
Query The Agent: Improving sample efficiency through epistemic uncertainty estimation5 Oct 2022 0 repositories listed
-
Using Deep Reinforcement Learning for mmWave Real-Time Scheduling4 Oct 2022 0 repositories listed
-
Federated Reinforcement Learning for Real-Time Electric Vehicle Charging and Discharging Control4 Oct 2022 0 repositories listed
-
Handling Sparse Rewards in Reinforcement Learning Using Model Predictive Control4 Oct 2022 0 repositories listed
-
Hyperbolic Deep Reinforcement Learning4 Oct 2022 0 repositories listed
-
Learning Dynamic Abstract Representations for Sample-Efficient Reinforcement Learning4 Oct 2022 0 repositories listed
-
Maximum-Likelihood Inverse Reinforcement Learning with Finite-Time Guarantees4 Oct 2022 0 repositories listed
-
CostNet: An End-to-End Framework for Goal-Directed Reinforcement Learning3 Oct 2022 0 repositories listed
-
Interpretable Option Discovery using Deep Q-Learning and Variational Autoencoders3 Oct 2022 0 repositories listed
-
MSRL: Distributed Reinforcement Learning with Dataflow Fragments3 Oct 2022 0 repositories listed
-
Near-Optimal Deployment Efficiency in Reward-Free Reinforcement Learning with Linear Function Approximation3 Oct 2022 0 repositories listed
-
Offline Reinforcement Learning with Differentiable Function Approximation is Provably Efficient3 Oct 2022 0 repositories listed
-
Policy Gradient for Reinforcement Learning with General Utilities3 Oct 2022 0 repositories listed
-
Square-root regret bounds for continuous-time episodic Markov decision processes3 Oct 2022 0 repositories listed
-
EUCLID: Towards Efficient Unsupervised Reinforcement Learning with Multi-choice Dynamics Model2 Oct 2022 0 repositories listed
-
Policy Gradients for Probabilistic Constrained Reinforcement Learning2 Oct 2022 0 repositories listed
-
Robust Bayesian optimization with reinforcement learned acquisition functions2 Oct 2022 0 repositories listed
-
Bayesian Q-learning With Imperfect Expert Demonstrations1 Oct 2022 0 repositories listed
-
Comparing BERT-based Reward Functions for Deep Reinforcement Learning in Machine Translation1 Oct 2022 0 repositories listed
-
Parsing Natural Language into Propositional and First-Order Logic with Dual Reinforcement Learning1 Oct 2022 0 repositories listed
-
Zero-Shot Policy Transfer with Disentangled Task Representation of Meta-Reinforcement Learning1 Oct 2022 0 repositories listed
-
A General Framework for Sample-Efficient Function Approximation in Reinforcement Learning30 Sep 2022 0 repositories listed
-
ASPiRe:Adaptive Skill Priors for Reinforcement Learning30 Sep 2022 0 repositories listed
-
Efficiently Learning Small Policies for Locomotion and Manipulation30 Sep 2022 0 repositories listed
-
Bounded Robustness in Reinforcement Learning via Lexicographic Objectives30 Sep 2022 0 repositories listed
-
Programmable Control of Ultrasound Swarmbots through Reinforcement Learning30 Sep 2022 0 repositories listed
-
RL-MD: A Novel Reinforcement Learning Approach for DNA Motif Discovery30 Sep 2022 0 repositories listed
-
Blessing from Human-AI Interaction: Super Reinforcement Learning in Confounded Environments29 Sep 2022 0 repositories listed
-
Contrastive Unsupervised Learning of World Model with Invariant Causal Features29 Sep 2022 0 repositories listed
-
Ensemble Reinforcement Learning in Continuous Spaces -- A Hierarchical Multi-Step Approach for Policy Training29 Sep 2022 0 repositories listed
-
How Does Return Distribution in Distributional Reinforcement Learning Help Optimization?29 Sep 2022 0 repositories listed
-
Learning Parsimonious Dynamics for Generalization in Reinforcement Learning29 Sep 2022 0 repositories listed
-
Online Weighted Q-Ensembles for Reduced Hyperparameter Tuning in Reinforcement Learning29 Sep 2022 0 repositories listed
-
Reinforcement Learning Algorithms: An Overview and Classification29 Sep 2022 0 repositories listed
-
Argumentative Reward Learning: Reasoning About Human Preferences28 Sep 2022 0 repositories listed
-
Disentangling Transfer in Continual Reinforcement Learning28 Sep 2022 0 repositories listed
-
FIRE: A Failure-Adaptive Reinforcement Learning Framework for Edge Computing Migrations28 Sep 2022 0 repositories listed
-
Guiding Safe Exploration with Weakest Preconditions28 Sep 2022 0 repositories listed
-
Generalization in Deep Reinforcement Learning for Robotic Navigation by Reward Shaping28 Sep 2022 0 repositories listed
-
Predictive Crypto-Asset Automated Market Making Architecture for Decentralized Finance using Deep Reinforcement Learning28 Sep 2022 0 repositories listed
-
DCE: Offline Reinforcement Learning With Double Conservative Estimates27 Sep 2022 0 repositories listed
-
Reinforcement Learning with Non-Exponential Discounting27 Sep 2022 0 repositories listed
-
Safe Reinforcement Learning of Dynamic High-Dimensional Robotic Tasks: Navigation, Manipulation, Interaction27 Sep 2022 0 repositories listed
-
DEFT: Diverse Ensembles for Fast Transfer in Reinforcement Learning26 Sep 2022 0 repositories listed
-
Delayed Geometric Discounts: An Alternative Criterion for Reinforcement Learning26 Sep 2022 0 repositories listed
-
Overcoming Referential Ambiguity in Language-Guided Goal-Conditioned Reinforcement Learning26 Sep 2022 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.