Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 71
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 71 of 152: papers 7,001 to 7,100 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Policy Optimization for Continuous Reinforcement Learning30 May 2023 0 repositories listed
-
RL + Model-based Control: Using On-demand Optimal Control to Learn Versatile Legged Locomotion29 May 2023 0 repositories listed
-
RLAD: Reinforcement Learning from Pixels for Autonomous Driving in Urban Environments29 May 2023 0 repositories listed
-
Towards a Better Understanding of Representation Dynamics under TD-learning29 May 2023 0 repositories listed
-
Potential-based Credit Assignment for Cooperative RL-based Testing of Autonomous Vehicles28 May 2023 0 repositories listed
-
Distributional Reinforcement Learning with Dual Expectile-Quantile Regression26 May 2023 0 repositories listed
-
Emergent Agentic Transformer from Chain of Hindsight Experience26 May 2023 0 repositories listed
-
Learning Interpretable Models of Aircraft Handling Behaviour by Reinforcement Learning from Human Feedback26 May 2023 0 repositories listed
-
Policy Synthesis and Reinforcement Learning for Discounted LTL26 May 2023 0 repositories listed
-
Reinforcement Learning with Simple Sequence Priors26 May 2023 0 repositories listed
-
The Curious Price of Distributional Robustness in Reinforcement Learning with a Generative Model26 May 2023 0 repositories listed
-
Deterministic policy gradient based optimal control with probabilistic constraints25 May 2023 0 repositories listed
-
A Mini Review on the utilization of Reinforcement Learning with OPC UA24 May 2023 0 repositories listed
-
Control invariant set enhanced safe reinforcement learning: improved sampling efficiency, guaranteed stability and robustness24 May 2023 0 repositories listed
-
Deep Reinforcement Learning with Plasticity Injection24 May 2023 0 repositories listed
-
Matrix Estimation for Offline Reinforcement Learning with Low-Rank Structure24 May 2023 0 repositories listed
-
Combining Multi-Objective Bayesian Optimization with Reinforcement Learning for TinyML23 May 2023 0 repositories listed
-
ChemGymRL: An Interactive Framework for Reinforcement Learning for Digital Chemistry23 May 2023 0 repositories listed
-
Constrained Proximal Policy Optimization23 May 2023 0 repositories listed
-
Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning23 May 2023 0 repositories listed
-
Adaptive action supervision in reinforcement learning from real-world multi-agent demonstrations22 May 2023 0 repositories listed
-
Lagrangian-based online safe reinforcement learning for state-constrained systems22 May 2023 0 repositories listed
-
INVICTUS: Optimizing Boolean Logic Circuit Synthesis via Synergistic Learning and Search22 May 2023 0 repositories listed
-
Offline Primal-Dual Reinforcement Learning for Linear MDPs22 May 2023 0 repositories listed
-
Towards Optimal Energy Management Strategy for Hybrid Electric Vehicle with Reinforcement Learning21 May 2023 0 repositories listed
-
Model-based adaptation for sample efficient transfer in reinforcement learning control of parameter-varying systems20 May 2023 0 repositories listed
-
Understanding the World to Solve Social Dilemmas Using Multi-Agent Reinforcement Learning19 May 2023 0 repositories listed
-
Optimistic Natural Policy Gradient: a Simple Efficient Policy Optimization Framework for Online RL18 May 2023 0 repositories listed
-
The Blessing of Heterogeneity in Federated Q-Learning: Linear Speedup and Beyond18 May 2023 0 repositories listed
-
Reward-agnostic Fine-tuning: Provable Statistical Benefits of Hybrid Reinforcement Learning17 May 2023 0 repositories listed
-
Coagent Networks: Generalized and Scaled16 May 2023 0 repositories listed
-
Cooperation Is All You Need16 May 2023 0 repositories listed
-
Reinforcement Learning for Safe Robot Control using Control Lyapunov Barrier Functions16 May 2023 0 repositories listed
-
A Theoretical Analysis of Optimistic Proximal Policy Optimization in Linear Markov Decision Processes15 May 2023 0 repositories listed
-
Attention-based QoE-aware Digital Twin Empowered Edge Computing for Immersive Virtual Reality15 May 2023 0 repositories listed
-
Horizon-free Reinforcement Learning in Adversarial Linear Mixture MDPs15 May 2023 0 repositories listed
-
Task-Oriented Communication Design at Scale15 May 2023 0 repositories listed
-
Uniform-PAC Guarantees for Model-Based RL with Bounded Eluder Dimension15 May 2023 0 repositories listed
-
Delay-Adapted Policy Optimization and Improved Regret for Adversarial MDP with Delayed Bandit Feedback13 May 2023 0 repositories listed
-
12 May 2023 0 repositories listed Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm11 May 2023 0 repositories listed
-
Optimizing Memory Mapping Using Deep Reinforcement Learning11 May 2023 0 repositories listed
-
An Option-Dependent Analysis of Regret Minimization Algorithms in Finite-Horizon Semi-Markov Decision Processes10 May 2023 0 repositories listed
-
Discovery of Optimal Quantum Error Correcting Codes via Reinforcement Learning10 May 2023 0 repositories listed
-
Supplementing Gradient-Based Reinforcement Learning with Simple Evolutionary Ideas10 May 2023 0 repositories listed
-
Assessment of Reinforcement Learning Algorithms for Nuclear Power Plant Fuel Optimization9 May 2023 0 repositories listed
-
9 May 2023 0 repositories listed
-
RLocator: Reinforcement Learning for Bug Localization9 May 2023 0 repositories listed
-
Knowledge-enhanced Agents for Interactive Text Games8 May 2023 0 repositories listed
-
Truncating Trajectories in Monte Carlo Reinforcement Learning7 May 2023 0 repositories listed
-
Replicating Complex Dialogue Policy of Humans via Offline Imitation Learning with Supervised Regularization6 May 2023 0 repositories listed
-
How to Use Reinforcement Learning to Facilitate Future Electricity Market Design? Part 1: A Paradigmatic Theory4 May 2023 0 repositories listed
-
How to Use Reinforcement Learning to Facilitate Future Electricity Market Design? Part 2: Method and Applications4 May 2023 0 repositories listed
-
Rethinking Population-assisted Off-policy Reinforcement Learning4 May 2023 0 repositories listed
-
Gym-preCICE: Reinforcement Learning Environments for Active Flow Control3 May 2023 0 repositories listed
-
Validation of massively-parallel adaptive testing using dynamic control matching2 May 2023 0 repositories listed
-
Joint Learning of Policy with Unknown Temporal Constraints for Safe Reinforcement Learning30 Apr 2023 0 repositories listed
-
A Transfer Learning Approach to Minimize Reinforcement Learning Risks in Energy Optimization for Smart Buildings30 Apr 2023 0 repositories listed
-
A Federated Reinforcement Learning Framework for Link Activation in Multi-link Wi-Fi Networks28 Apr 2023 0 repositories listed
-
One-Step Distributional Reinforcement Learning27 Apr 2023 0 repositories listed
-
Can Agents Run Relay Race with Strangers? Generalization of RL to Out-of-Distribution Trajectories26 Apr 2023 0 repositories listed
-
Multi-criteria Hardware Trojan Detection: A Reinforcement Learning Approach26 Apr 2023 0 repositories listed
-
Reinforcement Learning with Partial Parametric Model Knowledge26 Apr 2023 0 repositories listed
-
A Closer Look at Reward Decomposition for High-Level Robotic Explanations25 Apr 2023 0 repositories listed
-
Loss- and Reward-Weighting for Efficient Distributed Reinforcement Learning25 Apr 2023 0 repositories listed
-
Model Extraction Attacks Against Reinforcement Learning Based Controllers25 Apr 2023 0 repositories listed
-
What can online reinforcement learning with function approximation benefit from general coverage conditions?25 Apr 2023 0 repositories listed
-
On Dynamic Programming Decompositions of Static Risk Measures in Markov Decision Processes24 Apr 2023 0 repositories listed
-
Policy Resilience to Environment Poisoning Attacks on Reinforcement Learning24 Apr 2023 0 repositories listed
-
Reinforcement Learning with Knowledge Representation and Reasoning: A Brief Survey24 Apr 2023 0 repositories listed
-
A Cubic-regularized Policy Newton Algorithm for Reinforcement Learning21 Apr 2023 0 repositories listed
-
A Review of Symbolic, Subsymbolic and Hybrid Methods for Sequential Decision Making20 Apr 2023 0 repositories listed
-
End-to-End Policy Gradient Method for POMDPs and Explainable Agents19 Apr 2023 0 repositories listed
-
FastRLAP: A System for Learning High-Speed Driving via Deep RL and Autonomous Practicing19 Apr 2023 0 repositories listed
-
Learning and Adapting Agile Locomotion Skills by Transferring Experience19 Apr 2023 0 repositories listed
-
Cooperative Multi-Agent Reinforcement Learning for Inventory Management18 Apr 2023 0 repositories listed
-
Feasible Policy Iteration for Safe Reinforcement Learning18 Apr 2023 0 repositories listed
-
Provably Feedback-Efficient Reinforcement Learning via Active Reward Learning18 Apr 2023 0 repositories listed
-
An adaptive safety layer with hard constraints for safe reinforcement learning in multi-energy management systems18 Apr 2023 0 repositories listed
-
MDDL: A Framework for Reinforcement Learning-based Position Allocation in Multi-Channel Feed17 Apr 2023 0 repositories listed
-
Car-Following Models: A Multidisciplinary Review14 Apr 2023 0 repositories listed
-
Bandit-Based Policy Invariant Explicit Shaping for Incorporating External Advice in Reinforcement Learning14 Apr 2023 0 repositories listed
-
Minimax-Optimal Reward-Agnostic Exploration in Reinforcement Learning14 Apr 2023 0 repositories listed
-
Exploring the Noise Resilience of Successor Features and Predecessor Features Algorithms in One and Two-Dimensional Environments14 Apr 2023 0 repositories listed
-
Towards Controllable Diffusion Models via Reward-Guided Exploration14 Apr 2023 0 repositories listed
-
Model-based Dynamic Shielding for Safe and Efficient Multi-Agent Reinforcement Learning13 Apr 2023 0 repositories listed
-
Facilitating Sim-to-real by Intrinsic Stochasticity of Real-Time Simulation in Reinforcement Learning for Robot Manipulation12 Apr 2023 0 repositories listed
-
Human-Robot Skill Transfer with Enhanced Compliance via Dynamic Movement Primitives12 Apr 2023 0 repositories listed
-
Multi-agent Policy Reciprocity with Theoretical Guarantee12 Apr 2023 0 repositories listed
-
Control invariant set enhanced reinforcement learning for process control: improved sampling efficiency and guaranteed stability11 Apr 2023 0 repositories listed
-
Optimal Interpretability-Performance Trade-off of Classification Trees with Black-Box Reinforcement Learning11 Apr 2023 0 repositories listed
-
For Pre-Trained Vision Models in Motor Control, Not All Policy Learning Methods are Created Equal10 Apr 2023 0 repositories listed
-
Learning a Universal Human Prior for Dexterous Manipulation from Human Preference10 Apr 2023 0 repositories listed
-
AI-Driven Resource Allocation in Optical Wireless Communication Systems8 Apr 2023 0 repositories listed
-
DREAM: Adaptive Reinforcement Learning based on Attention Mechanism for Temporal Knowledge Graph Reasoning8 Apr 2023 0 repositories listed
-
Evolving Reinforcement Learning Environment to Minimize Learner's Achievable Reward: An Application on Hardening Active Directory Systems8 Apr 2023 0 repositories listed
-
Continuous Input Embedding Size Search For Recommender Systems7 Apr 2023 0 repositories listed
-
Persuading to Prepare for Quitting Smoking with a Virtual Coach: Using States and User Characteristics to Predict Behavior5 Apr 2023 0 repositories listed
-
A Multiagent CyberBattleSim for RL Cyber Operation Agents3 Apr 2023 0 repositories listed
-
A Tutorial Introduction to Reinforcement Learning3 Apr 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.