Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 72
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 72 of 152: papers 7,101 to 7,200 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Enabling A Network AI Gym for Autonomous Cyber Agents3 Apr 2023 0 repositories listed
-
Quantitative Trading using Deep Q Learning3 Apr 2023 0 repositories listed
-
Unified Emulation-Simulation Training Environment for Autonomous Cyber Agents3 Apr 2023 0 repositories listed
-
Risk-Sensitive and Robust Model-Based Reinforcement Learning and Planning2 Apr 2023 0 repositories listed
-
Mastering Pair Trading with Risk-Aware Recurrent Reinforcement Learning1 Apr 2023 0 repositories listed
-
Restarted Bayesian Online Change-point Detection for Non-Stationary Markov Decision Processes1 Apr 2023 0 repositories listed
-
Accelerating exploration and representation learning with offline pre-training31 Mar 2023 0 repositories listed
-
Understanding Reinforcement Learning Algorithms: The Progress from Basic Q-learning to Proximal Policy Optimization31 Mar 2023 0 repositories listed
-
Finetuning from Offline Reinforcement Learning: Challenges, Trade-offs and Practical Solutions30 Mar 2023 0 repositories listed
-
Learning in Factored Domains with Information-Constrained Visual Representations30 Mar 2023 0 repositories listed
-
On the Analysis of Computational Delays in Reinforcement Learning-based Rate Adaptation Algorithms30 Mar 2023 0 repositories listed
-
When Learning Is Out of Reach, Reset: Generalization in Autonomous Visuomotor Reinforcement Learning30 Mar 2023 0 repositories listed
-
Does Sparsity Help in Learning Misspecified Linear Bandits?29 Mar 2023 0 repositories listed
-
Skill Reinforcement Learning and Planning for Open-World Long-Horizon Tasks29 Mar 2023 0 repositories listed
-
Planning with Sequence Models through Iterative Energy Minimization28 Mar 2023 0 repositories listed
-
On-line reinforcement learning for optimization of real-life energy trading strategy28 Mar 2023 0 repositories listed
-
Bi-Manual Block Assembly via Sim-to-Real Reinforcement Learning27 Mar 2023 0 repositories listed
-
Multi-Flow Transmission in Wireless Interference Networks: A Convergent Graph Learning Approach27 Mar 2023 0 repositories listed
-
Robust Risk-Aware Option Hedging27 Mar 2023 0 repositories listed
-
Control of synaptic plasticity via the fusion of reinforcement learning and unsupervised learning in neural networks26 Mar 2023 0 repositories listed
-
Learning to Operate in Open Worlds by Adapting Planning Models24 Mar 2023 0 repositories listed
-
A Hierarchical Hybrid Learning Framework for Multi-agent Trajectory Prediction22 Mar 2023 0 repositories listed
-
Adaptive Road Configurations for Improved Autonomous Vehicle-Pedestrian Interactions using Reinforcement Learning22 Mar 2023 0 repositories listed
-
Communication Load Balancing via Efficient Inverse Reinforcement Learning22 Mar 2023 0 repositories listed
-
Deep RL with Hierarchical Action Exploration for Dialogue Generation22 Mar 2023 0 repositories listed
-
Policy Reuse for Communication Load Balancing in Unseen Traffic Scenarios22 Mar 2023 0 repositories listed
-
Synthetic Health-related Longitudinal Data with Mixed-type Variables Generated using Diffusion Models22 Mar 2023 0 repositories listed
-
Beam Management Driven by Radio Environment Maps in O-RAN Architecture21 Mar 2023 0 repositories listed
-
A Survey of Demonstration Learning20 Mar 2023 0 repositories listed
-
Bridging Imitation and Online Reinforcement Learning: An Optimistic Tale20 Mar 2023 0 repositories listed
-
Deceptive Reinforcement Learning in Model-Free Domains20 Mar 2023 0 repositories listed
-
Improved Sample Complexity for Reward-free Reinforcement Learning under Low-rank MDPs20 Mar 2023 0 repositories listed
-
Active hypothesis testing in unknown environments using recurrent neural networks and model free reinforcement learning19 Mar 2023 0 repositories listed
-
Boundary-aware Supervoxel-level Iteratively Refined Interactive 3D Image Segmentation with Multi-agent Reinforcement Learning19 Mar 2023 0 repositories listed
-
Cheap Talk Discovery and Utilization in Multi-Agent Reinforcement Learning19 Mar 2023 0 repositories listed
-
Multi-modal reward for visual relationships-based image captioning19 Mar 2023 0 repositories listed
-
Hybrid Systems Neural Control with Region-of-Attraction Planner18 Mar 2023 0 repositories listed
-
Interpretable Reinforcement Learning via Neural Additive Models for Inventory Management18 Mar 2023 0 repositories listed
-
A Data-Driven Model-Reference Adaptive Control Approach Based on Reinforcement Learning17 Mar 2023 0 repositories listed
-
A New Policy Iteration Algorithm For Reinforcement Learning in Zero-Sum Markov Games17 Mar 2023 0 repositories listed
-
Comparing NARS and Reinforcement Learning: An Analysis of ONA and Q-Learning Algorithms17 Mar 2023 0 repositories listed
-
Measurement Optimization under Uncertainty using Deep Reinforcement Learning17 Mar 2023 0 repositories listed
-
Optimal Horizon-Free Reward-Free Exploration for Linear Mixture MDPs17 Mar 2023 0 repositories listed
-
Towards Real-World Applications of Personalized Anesthesia Using Policy Constraint Q Learning for Propofol Infusion Control17 Mar 2023 0 repositories listed
-
Efficient Learning of High Level Plans from Play16 Mar 2023 0 repositories listed
-
Goal-conditioned Offline Reinforcement Learning through State Space Partitioning16 Mar 2023 0 repositories listed
-
Learning Rewards to Optimize Global Performance Metrics in Deep Reinforcement Learning16 Mar 2023 0 repositories listed
-
Online Reinforcement Learning in Periodic MDP16 Mar 2023 0 repositories listed
-
Psychotherapy AI Companion with Reinforcement Learning Recommendations and Interpretable Policy Dynamics16 Mar 2023 0 repositories listed
-
Recommending the optimal policy by learning to act from temporal data16 Mar 2023 0 repositories listed
-
Reinforcement Learning for Omega-Regular Specifications on Continuous-Time MDP16 Mar 2023 0 repositories listed
-
Self-Inspection Method of Unmanned Aerial Vehicles in Power Plants Using Deep Q-Network Reinforcement Learning16 Mar 2023 0 repositories listed
-
SVDE: Scalable Value-Decomposition Exploration for Cooperative Multi-Agent Reinforcement Learning16 Mar 2023 0 repositories listed
-
Latent-Conditioned Policy Gradient for Multi-Objective Deep Reinforcement Learning15 Mar 2023 0 repositories listed
-
Muti-Agent Proximal Policy Optimization For Data Freshness in UAV-assisted Networks15 Mar 2023 0 repositories listed
-
On the Benefits of Leveraging Structural Information in Planning Over the Learned Model15 Mar 2023 0 repositories listed
-
Real-Time Measurement-Driven Reinforcement Learning Control Approach for Uncertain Nonlinear Systems15 Mar 2023 0 repositories listed
-
Replay Buffer with Local Forgetting for Adapting to Local Environment Changes in Deep Model-Based Reinforcement Learning15 Mar 2023 0 repositories listed
-
Smoothed Q-learning15 Mar 2023 0 repositories listed
-
Optimizing Trading Strategies in Quantitative Markets using Multi-Agent Reinforcement Learning15 Mar 2023 0 repositories listed
-
Adaptive Policy Learning for Offline-to-Online Reinforcement Learning14 Mar 2023 0 repositories listed
-
Actor-Critic learning for mean-field control in continuous time13 Mar 2023 0 repositories listed
-
Deploying Offline Reinforcement Learning with Human Feedback13 Mar 2023 0 repositories listed
-
Loss of Plasticity in Continual Deep Reinforcement Learning13 Mar 2023 0 repositories listed
-
Path Planning using Reinforcement Learning: A Policy Iteration Approach13 Mar 2023 0 repositories listed
-
Reinforcement Learning-based Wavefront Sensorless Adaptive Optics Approaches for Satellite-to-Ground Laser Communication13 Mar 2023 0 repositories listed
-
Visual-Policy Learning through Multi-Camera View to Single-Camera View Knowledge Distillation for Robot Manipulation Tasks13 Mar 2023 0 repositories listed
-
Behavioral Differences is the Key of Ad-hoc Team Cooperation in Multiplayer Games Hanabi12 Mar 2023 0 repositories listed
-
The tree reconstruction game: phylogenetic reconstruction using reinforcement learning12 Mar 2023 0 repositories listed
-
Provably Efficient Model-Free Algorithms for Non-stationary CMDPs10 Mar 2023 0 repositories listed
-
Understanding the Synergies between Quality-Diversity and Deep Reinforcement Learning10 Mar 2023 0 repositories listed
-
A Framework for History-Aware Hyperparameter Optimisation in Reinforcement Learning9 Mar 2023 0 repositories listed
-
Beware of Instantaneous Dependence in Reinforcement Learning9 Mar 2023 0 repositories listed
-
Computably Continuous Reinforcement-Learning Objectives are PAC-learnable9 Mar 2023 0 repositories listed
-
Conceptual Reinforcement Learning for Language-Conditioned Tasks9 Mar 2023 0 repositories listed
-
Exploiting Contextual Structure to Generate Useful Auxiliary Tasks9 Mar 2023 0 repositories listed
-
GOATS: Goal Sampling Adaptation for Scooping with Curriculum Reinforcement Learning9 Mar 2023 0 repositories listed
-
Power and Interference Control for VLC-Based UDN: A Reinforcement Learning Approach9 Mar 2023 0 repositories listed
-
Real-time scheduling of renewable power systems through planning-based reinforcement learning9 Mar 2023 0 repositories listed
-
Recent Advances of Deep Robotic Affordance Learning: A Reinforcement Learning Perspective9 Mar 2023 0 repositories listed
-
Task Aware Dreamer for Task Generalization in Reinforcement Learning9 Mar 2023 0 repositories listed
-
Variance-aware robust reinforcement learning with linear function approximation under heavy-tailed rewards9 Mar 2023 0 repositories listed
-
Using Memory-Based Learning to Solve Tasks with State-Action Constraints8 Mar 2023 0 repositories listed
-
adaPARL: Adaptive Privacy-Aware Reinforcement Learning for Sequential-Decision Making Human-in-the-Loop Systems7 Mar 2023 0 repositories listed
-
Decoupling Skill Learning from Robotic Control for Generalizable Object Manipulation7 Mar 2023 0 repositories listed
-
Deep Occupancy-Predictive Representations for Autonomous Driving7 Mar 2023 0 repositories listed
-
Domain Randomization for Robust, Affordable and Effective Closed-loop Control of Soft Robots7 Mar 2023 0 repositories listed
-
Environment Transformer and Policy Optimization for Model-Based Offline Reinforcement Learning7 Mar 2023 0 repositories listed
-
Evolutionary Reinforcement Learning: A Survey7 Mar 2023 0 repositories listed
-
Graph Decision Transformer7 Mar 2023 0 repositories listed
-
On the Sample Complexity of Vanilla Model-Based Offline Reinforcement Learning with Dependent Samples7 Mar 2023 0 repositories listed
-
Efficient Skill Acquisition for Complex Manipulation Tasks in Obstructed Environments6 Mar 2023 0 repositories listed
-
MAESTRO: Open-Ended Environment Design for Multi-Agent Reinforcement Learning6 Mar 2023 0 repositories listed
-
Perspectives on the Social Impacts of Reinforcement Learning with Human Feedback6 Mar 2023 0 repositories listed
-
Reinforcement Learning Based Self-play and State Stacking Techniques for Noisy Air Combat Environment6 Mar 2023 0 repositories listed
-
Dexterous In-hand Manipulation by Guiding Exploration with Simple Sub-skill Controllers6 Mar 2023 0 repositories listed
-
Ensemble Reinforcement Learning: A Survey5 Mar 2023 0 repositories listed
-
Local Environment Poisoning Attacks on Federated Reinforcement Learning5 Mar 2023 0 repositories listed
-
Sparsity-Aware Intelligent Massive Random Access Control in Open RAN: A Reinforcement Learning Based Approach5 Mar 2023 0 repositories listed
-
Double A3C: Deep Reinforcement Learning on OpenAI Gym Games4 Mar 2023 0 repositories listed