Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 79
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 79 of 152: papers 7,801 to 7,900 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
MetaEMS: A Meta Reinforcement Learning-based Control Framework for Building Energy Management System23 Oct 2022 0 repositories listed
-
Attitude Control of Highly Maneuverable Aircraft Using an Improved Q-learning22 Oct 2022 0 repositories listed
-
Faster and more diverse de novo molecular optimization with double-loop reinforcement learning using augmented SMILES22 Oct 2022 0 repositories listed
-
Probing Transfer in Deep Reinforcement Learning without Task Engineering22 Oct 2022 0 repositories listed
-
Continual Vision-based Reinforcement Learning with Group Symmetries21 Oct 2022 0 repositories listed
-
Redefining Counterfactual Explanations for Reinforcement Learning: Overview, Challenges and Opportunities21 Oct 2022 0 repositories listed
-
Deep Reinforcement Learning for Inverse Inorganic Materials Design21 Oct 2022 0 repositories listed
-
Deep Reinforcement Learning for Stabilization of Large-scale Probabilistic Boolean Networks21 Oct 2022 0 repositories listed
-
Group Distributionally Robust Reinforcement Learning with Hierarchical Latent Variables21 Oct 2022 0 repositories listed
-
Implicit Offline Reinforcement Learning via Supervised Learning21 Oct 2022 0 repositories listed
-
Integrating Policy Summaries with Reward Decomposition for Explaining Reinforcement Learning Agents21 Oct 2022 0 repositories listed
-
On the connection between Bregman divergence and value in regularized Markov decision processes21 Oct 2022 0 repositories listed
-
Epistemic Monte Carlo Tree Search21 Oct 2022 0 repositories listed
-
Towards Quantum-Enabled 6G Slicing21 Oct 2022 0 repositories listed
-
Fine-Grained Session Recommendations in E-commerce using Deep Reinforcement Learning20 Oct 2022 0 repositories listed
-
Horizon-Free and Variance-Dependent Reinforcement Learning for Latent Markov Decision Processes20 Oct 2022 0 repositories listed
-
Robust Imitation via Mirror Descent Inverse Reinforcement Learning20 Oct 2022 0 repositories listed
-
Safe Policy Improvement in Constrained Markov Decision Processes20 Oct 2022 0 repositories listed
-
A Reinforcement Learning Approach in Multi-Phase Second-Price Auction Design19 Oct 2022 0 repositories listed
-
Hierarchical Reinforcement Learning for Furniture Layout in Virtual Indoor Scenes19 Oct 2022 0 repositories listed
-
Integrated Decision and Control for High-Level Automated Vehicles by Mixed Policy Gradient and Its Experiment Verification19 Oct 2022 0 repositories listed
-
On the Power of Pre-training for Generalization in RL: Provable Benefits and Hardness19 Oct 2022 0 repositories listed
-
Oracles & Followers: Stackelberg Equilibria in Deep Multi-Agent Reinforcement Learning19 Oct 2022 0 repositories listed
-
Palm up: Playing in the Latent Manifold for Unsupervised Pretraining19 Oct 2022 0 repositories listed
-
Provably Safe Reinforcement Learning via Action Projection using Reachability Analysis and Polynomial Zonotopes19 Oct 2022 0 repositories listed
-
Robot Navigation with Reinforcement Learned Path Generation and Fine-Tuned Motion Control19 Oct 2022 0 repositories listed
-
Robotic Table Wiping via Reinforcement Learning and Whole-body Trajectory Optimization19 Oct 2022 0 repositories listed
-
Scaling Laws for Reward Model Overoptimization19 Oct 2022 0 repositories listed
-
RPM: Generalizable Behaviors for Multi-Agent Reinforcement Learning18 Oct 2022 0 repositories listed
-
Unpacking Reward Shaping: Understanding the Benefits of Reward Engineering on Sample Complexity18 Oct 2022 0 repositories listed
-
Boosting Offline Reinforcement Learning via Data Rebalancing17 Oct 2022 0 repositories listed
-
Model Predictive Control via On-Policy Imitation Learning17 Oct 2022 0 repositories listed
-
PTDE: Personalized Training with Distilled Execution for Multi-Agent Reinforcement Learning17 Oct 2022 0 repositories listed
-
You Only Live Once: Single-Life Reinforcement Learning17 Oct 2022 0 repositories listed
-
Data-Efficient Pipeline for Offline Reinforcement Learning with Limited Data16 Oct 2022 0 repositories listed
-
Entropy Regularized Reinforcement Learning with Cascading Networks16 Oct 2022 0 repositories listed
-
The Impact of Task Underspecification in Evaluating Deep Reinforcement Learning16 Oct 2022 0 repositories listed
-
Towards an Interpretable Hierarchical Agent Framework using Semantic Goals16 Oct 2022 0 repositories listed
-
A Scalable Reinforcement Learning Approach for Attack Allocation in Swarm to Swarm Engagement Problems15 Oct 2022 0 repositories listed
-
DyFEn: Agent-Based Fee Setting in Payment Channel Networks15 Oct 2022 0 repositories listed
-
Near-Optimal Regret Bounds for Multi-batch Reinforcement Learning15 Oct 2022 0 repositories listed
-
PI-QT-Opt: Predictive Information Improves Multi-Task Robotic Reinforcement Learning at Scale15 Oct 2022 0 repositories listed
-
Reinforcement Learning for ConnectX15 Oct 2022 0 repositories listed
-
Revisiting the Roles of "Text" in Text Games15 Oct 2022 0 repositories listed
-
A Reinforcement Learning Approach to Estimating Long-term Treatment Effects14 Oct 2022 0 repositories listed
-
A Scalable Finite Difference Method for Deep Reinforcement Learning14 Oct 2022 0 repositories listed
-
Query Rewriting for Effective Misinformation Discovery14 Oct 2022 0 repositories listed
-
Adaptive patch foraging in deep reinforcement learning agents14 Oct 2022 0 repositories listed
-
Multi-trainer Interactive Reinforcement Learning System14 Oct 2022 0 repositories listed
-
Robust Preference Learning for Storytelling via Contrastive Reinforcement Learning14 Oct 2022 0 repositories listed
-
A Concise Introduction to Reinforcement Learning in Robotics13 Oct 2022 0 repositories listed
-
Causality-driven Hierarchical Structure Discovery for Reinforcement Learning13 Oct 2022 0 repositories listed
-
Deep reinforcement learning for automatic run-time adaptation of UWB PHY radio settings13 Oct 2022 0 repositories listed
-
Dissipative residual layers for unsupervised implicit parameterization of data manifolds13 Oct 2022 0 repositories listed
-
Efficient circuit implementation for coined quantum walks on binary trees and application to reinforcement learning13 Oct 2022 0 repositories listed
-
Object-Category Aware Reinforcement Learning13 Oct 2022 0 repositories listed
-
Observed Adversaries in Deep Reinforcement Learning13 Oct 2022 0 repositories listed
-
Optimal Control of Material Micro-Structures13 Oct 2022 0 repositories listed
-
Output Feedback Adaptive Optimal Control of Affine Nonlinear systems with a Linear Measurement Model13 Oct 2022 0 repositories listed
-
Personalized Federated Hypernetworks for Privacy Preservation in Multi-Task Reinforcement Learning13 Oct 2022 0 repositories listed
-
Policy Gradient With Serial Markov Chain Reasoning13 Oct 2022 0 repositories listed
-
Reinforcement Learning with Unbiased Policy Evaluation and Linear Function Approximation13 Oct 2022 0 repositories listed
-
Towards Multi-Agent Reinforcement Learning driven Over-The-Counter Market Simulations13 Oct 2022 0 repositories listed
-
DQLAP: Deep Q-Learning Recommender Algorithm with Update Policy for a Real Steam Turbine System12 Oct 2022 0 repositories listed
-
Explaining Online Reinforcement Learning Decisions of Self-Adaptive Systems12 Oct 2022 0 repositories listed
-
Real World Offline Reinforcement Learning with Realistic Data Source12 Oct 2022 0 repositories listed
-
Reinforcement Learning with Automated Auxiliary Loss Search12 Oct 2022 0 repositories listed
-
Smooth Trajectory Collision Avoidance through Deep Reinforcement Learning12 Oct 2022 0 repositories listed
-
Broad-persistent Advice for Interactive Reinforcement Learning Scenarios11 Oct 2022 0 repositories listed
-
Edge-Cloud Cooperation for DNN Inference via Reinforcement Learning and Supervised Learning11 Oct 2022 0 repositories listed
-
Multi-User Reinforcement Learning with Low Rank Rewards11 Oct 2022 0 repositories listed
-
Regret Bounds for Risk-Sensitive Reinforcement Learning11 Oct 2022 0 repositories listed
-
The Role of Exploration for Task Transfer in Reinforcement Learning11 Oct 2022 0 repositories listed
-
Creating a Dynamic Quadrupedal Robotic Goalkeeper with Reinforcement Learning10 Oct 2022 0 repositories listed
-
Long N-step Surrogate Stage Reward to Reduce Variances of Deep Reinforcement Learning in Complex Problems10 Oct 2022 0 repositories listed
-
Simulating Coverage Path Planning with Roomba10 Oct 2022 0 repositories listed
-
Towards a Theoretical Foundation of Policy Optimization for Learning Control Policies10 Oct 2022 0 repositories listed
-
Equivalence of Optimality Criteria for Markov Decision Process and Model Predictive Control9 Oct 2022 0 repositories listed
-
The Role of Coverage in Online Reinforcement Learning9 Oct 2022 0 repositories listed
-
Cognitive Models as Simulators: The Case of Moral Decision-Making8 Oct 2022 0 repositories listed
-
Dynamically meeting performance objectives for multiple services on a service mesh8 Oct 2022 0 repositories listed
-
Advice Conformance Verification by Reinforcement Learning agents for Human-in-the-Loop7 Oct 2022 0 repositories listed
-
Algorithmic Trading Using Continuous Action Space Deep Reinforcement Learning7 Oct 2022 0 repositories listed
-
How to Enable Uncertainty Estimation in Proximal Policy Optimization7 Oct 2022 0 repositories listed
-
Large Language Models can Implement Policy Iteration7 Oct 2022 0 repositories listed
-
Multi-agent Deep Covering Skill Discovery7 Oct 2022 0 repositories listed
-
Reinforcement Learning Approach for Multi-Agent Flexible Scheduling Problems7 Oct 2022 0 repositories listed
-
Deep Inventory Management6 Oct 2022 0 repositories listed
-
Digital Human Interactive Recommendation Decision-Making Based on Reinforcement Learning6 Oct 2022 0 repositories listed
-
6 Oct 2022 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
Learning Algorithms for Intelligent Agents and Mechanisms6 Oct 2022 0 repositories listed
-
Low-Thrust Orbital Transfer using Dynamics-Agnostic Reinforcement Learning6 Oct 2022 0 repositories listed
-
Lyapunov Function Consistent Adaptive Network Signal Control with Back Pressure and Reinforcement Learning6 Oct 2022 0 repositories listed
-
Meta Reinforcement Learning for Optimal Design of Legged Robots6 Oct 2022 0 repositories listed
-
Reinforcement Learning with Large Action Spaces for Neural Machine Translation6 Oct 2022 0 repositories listed
-
A Novel Entropy-Maximizing TD3-based Reinforcement Learning for Automatic PID Tuning5 Oct 2022 0 repositories listed
-
Neural Distillation as a State Representation Bottleneck in Reinforcement Learning5 Oct 2022 0 repositories listed
-
On Neural Consolidation for Transfer in Reinforcement Learning5 Oct 2022 0 repositories listed
-
Query The Agent: Improving sample efficiency through epistemic uncertainty estimation5 Oct 2022 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.