Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 108
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 108 of 152: papers 10,701 to 10,800 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Reducing Conservativeness Oriented Offline Reinforcement Learning27 Feb 2021 0 repositories listed
-
Revisiting Peng's Q(λ) for Modern Reinforcement Learning27 Feb 2021 0 repositories listed
-
Low-Precision Reinforcement Learning: Running Soft Actor-Critic in Half Precision26 Feb 2021 0 repositories listed
-
Potential Impacts of Smart Homes on Human Behavior: A Reinforcement Learning Approach26 Feb 2021 0 repositories listed
-
Safe Distributional Reinforcement Learning26 Feb 2021 0 repositories listed
-
Adaptive Load Shedding for Grid Emergency Control via Deep Reinforcement Learning25 Feb 2021 0 repositories listed
-
Twin actor twin delayed deep deterministic policy gradient (TATD3) learning for batch process control25 Feb 2021 0 repositories listed
-
Bias-reduced Multi-step Hindsight Experience Replay for Efficient Multi-goal Reinforcement Learning25 Feb 2021 0 repositories listed
-
Emerging Trends in Federated Learning: From Model Fusion to Federated X Learning25 Feb 2021 0 repositories listed
-
Iterative Bounding MDPs: Learning Interpretable Policies via Non-Interpretable Methods25 Feb 2021 0 repositories listed
-
No-Regret Reinforcement Learning with Heavy-Tailed Rewards25 Feb 2021 0 repositories listed
-
On The Effect of Auxiliary Tasks on Representation Dynamics25 Feb 2021 0 repositories listed
-
Reinforcement learning approach for resource allocation in humanitarian logistics25 Feb 2021 0 repositories listed
-
Reinforcement Learning of Implicit and Explicit Control Flow in Instructions25 Feb 2021 0 repositories listed
-
Where to go next: Learning a Subgoal Recommendation Policy for Navigation Among Pedestrians25 Feb 2021 0 repositories listed
-
Annotating Motion Primitives for Simplifying Action Search in Reinforcement Learning24 Feb 2021 0 repositories listed
-
Combining Off and On-Policy Training in Model-Based Reinforcement Learning24 Feb 2021 0 repositories listed
-
Beyond Fine-Tuning: Transferring Behavior in Reinforcement Learning24 Feb 2021 0 repositories listed
-
Credit Assignment with Meta-Policy Gradient for Multi-Agent Reinforcement Learning24 Feb 2021 0 repositories listed
-
Deep Reinforcement Learning for Safe Landing Site Selection with Concurrent Consideration of Divert Maneuvers24 Feb 2021 0 repositories listed
-
Fast Approximate Solutions using Reinforcement Learning for Dynamic Capacitated Vehicle Routing with Time Windows24 Feb 2021 0 repositories listed
-
FIXAR: A Fixed-Point Deep Reinforcement Learning Platform with Quantization-Aware Training and Adaptive Parallelism24 Feb 2021 0 repositories listed
-
Learning Emergent Discrete Message Communication for Cooperative Reinforcement Learning24 Feb 2021 0 repositories listed
-
PFRL: Pose-Free Reinforcement Learning for 6D Pose Estimation24 Feb 2021 0 repositories listed
-
The Logical Options Framework24 Feb 2021 0 repositories listed
-
Towards Safe Continuing Task Reinforcement Learning24 Feb 2021 0 repositories listed
-
A Robotic Model of Hippocampal Reverse Replay for Reinforcement Learning23 Feb 2021 0 repositories listed
-
DeepThermal: Combustion Optimization for Thermal Power Generating Units Using Offline Reinforcement Learning23 Feb 2021 0 repositories listed
-
Differentiable Logic Machines23 Feb 2021 0 repositories listed
-
Honey, I Shrunk The Actor: A Case Study on Preserving Performance with Smaller Actors in Actor-Critic RL23 Feb 2021 0 repositories listed
-
Greedy-Step Off-Policy Reinforcement Learning23 Feb 2021 0 repositories listed
-
MUSBO: Model-based Uncertainty Regularized and Sample Efficient Batch Optimization for Deployment Constrained Reinforcement Learning23 Feb 2021 0 repositories listed
-
School of hard knocks: Curriculum analysis for Pommerman with a fixed computational budget23 Feb 2021 0 repositories listed
-
State Augmented Constrained Reinforcement Learning: Overcoming the Limitations of Learning with Rewards23 Feb 2021 0 repositories listed
-
A Novel Framework for Neural Architecture Search in the Hill Climbing Domain22 Feb 2021 0 repositories listed
-
Action Redundancy in Reinforcement Learning22 Feb 2021 0 repositories listed
-
Communication Efficient Parallel Reinforcement Learning22 Feb 2021 0 repositories listed
-
Deep Reinforcement Learning for Dynamic Spectrum Sharing of LTE and NR22 Feb 2021 0 repositories listed
-
Efficient Text-based Reinforcement Learning by Jointly Leveraging State and Commonsense Graph Representations22 Feb 2021 0 repositories listed
-
Escaping from Zero Gradient: Revisiting Action-Constrained Reinforcement Learning via Frank-Wolfe Policy Optimization22 Feb 2021 0 repositories listed
-
Uncertainty Estimation Using Riemannian Model Dynamics for Offline Reinforcement Learning22 Feb 2021 0 repositories listed
-
Provably Improved Context-Based Offline Meta-RL with Attention and Contrastive Learning22 Feb 2021 0 repositories listed
-
Improved Learning of Robot Manipulation Tasks via Tactile Intrinsic Motivation22 Feb 2021 0 repositories listed
-
Reinforcement Learning of the Prediction Horizon in Model Predictive Control22 Feb 2021 0 repositories listed
-
Return-Based Contrastive Representation Learning for Reinforcement Learning22 Feb 2021 0 repositories listed
-
SENTINEL: Taming Uncertainty with Ensemble-based Distributional Reinforcement Learning22 Feb 2021 0 repositories listed
-
Stratified Experience Replay: Correcting Multiplicity Bias in Off-Policy Reinforcement Learning22 Feb 2021 0 repositories listed
-
Learning Efficient Navigation in Vortical Flow Fields21 Feb 2021 0 repositories listed
-
Safe Reinforcement Learning Using Robust Action Governor21 Feb 2021 0 repositories listed
-
How To Train Your HERON20 Feb 2021 0 repositories listed
-
Importance of Environment Design in Reinforcement Learning: A Study of a Robotic Environment20 Feb 2021 0 repositories listed
-
A Reinforcement Learning Approach to Age of Information in Multi-User Networks with HARQ19 Feb 2021 0 repositories listed
-
Decentralized Deterministic Multi-Agent Reinforcement Learning19 Feb 2021 0 repositories listed
-
Instrumental Variable Value Iteration for Causal Offline Reinforcement Learning19 Feb 2021 0 repositories listed
-
Model-Invariant State Abstractions for Model-Based Reinforcement Learning19 Feb 2021 0 repositories listed
-
TacticZero: Learning to Prove Theorems from Scratch with Deep Reinforcement Learning19 Feb 2021 0 repositories listed
-
Efficient Reinforcement Learning in Resource Allocation Problems Through Permutation Invariant Multi-task Learning18 Feb 2021 0 repositories listed
-
Learning Memory-Dependent Continuous Control from Demonstrations18 Feb 2021 0 repositories listed
-
Privacy-Preserving Kickstarting Deep Reinforcement Learning with Privacy-Aware Learners18 Feb 2021 0 repositories listed
-
Reinforcement Learning for Beam Pattern Design in Millimeter Wave and Massive MIMO Systems18 Feb 2021 0 repositories listed
-
Reinforcement Learning for Datacenter Congestion Control18 Feb 2021 0 repositories listed
-
Smart Feasibility Pump: Reinforcement Learning for (Mixed) Integer Programming18 Feb 2021 0 repositories listed
-
Strategic bidding in freight transport using deep reinforcement learning18 Feb 2021 0 repositories listed
-
Near-optimal Policy Optimization Algorithms for Learning Adversarial Linear Mixture MDPs17 Feb 2021 0 repositories listed
-
On the Convergence and Sample Efficiency of Variance-Reduced Policy Gradient Method17 Feb 2021 0 repositories listed
-
Separated Proportional-Integral Lagrangian for Chance Constrained Reinforcement Learning17 Feb 2021 0 repositories listed
-
Efficient Scheduling of Data Augmentation for Deep Reinforcement Learning17 Feb 2021 0 repositories listed
-
Active Privacy-utility Trade-off Against a Hypothesis Testing Adversary16 Feb 2021 0 repositories listed
-
Improper Reinforcement Learning with Gradient-based Policy Optimization16 Feb 2021 0 repositories listed
-
Inverse Reinforcement Learning in a Continuous State Space with Formal Guarantees16 Feb 2021 0 repositories listed
-
IronMan: GNN-assisted Design Space Exploration in High-Level Synthesis via Reinforcement Learning16 Feb 2021 0 repositories listed
-
Model-based Meta Reinforcement Learning using Graph Structured Surrogate Models16 Feb 2021 0 repositories listed
-
Multi-Stage Transmission Line Flow Control Using Centralized and Decentralized Reinforcement Learning Agents16 Feb 2021 0 repositories listed
-
Quantifying the effects of environment and population diversity in multi-agent reinforcement learning16 Feb 2021 0 repositories listed
-
Reward Poisoning in Reinforcement Learning: Attacks Against Unknown Learners in Unknown Environments16 Feb 2021 0 repositories listed
-
RMIX: Learning Risk-Sensitive Policies for Cooperative Reinforcement Learning Agents16 Feb 2021 0 repositories listed
-
TradeR: Practical Deep Hierarchical Reinforcement Learning for Trade Execution16 Feb 2021 0 repositories listed
-
Training Larger Networks for Deep Reinforcement Learning16 Feb 2021 0 repositories listed
-
Transferring Domain Knowledge with an Adviser in Continuous Tasks16 Feb 2021 0 repositories listed
-
Cooperation and Reputation Dynamics with Reinforcement Learning15 Feb 2021 0 repositories listed
-
Distributionally-Constrained Policy Optimization via Unbalanced Optimal Transport15 Feb 2021 0 repositories listed
-
Learning from Demonstrations using Signal Temporal Logic15 Feb 2021 0 repositories listed
-
Seeing by haptic glance: reinforcement learning-based 3D object Recognition15 Feb 2021 0 repositories listed
-
Domain Adversarial Reinforcement Learning14 Feb 2021 0 repositories listed
-
Model-free Representation Learning and Exploration in Low-rank MDPs14 Feb 2021 0 repositories listed
-
Reinforcement Learning for IoT Security: A Comprehensive Survey14 Feb 2021 0 repositories listed
-
Reversible Action Design for Combinatorial Optimization with Reinforcement Learning14 Feb 2021 0 repositories listed
-
Sparse Attention Guided Dynamic Value Estimation for Single-Task Multi-Scene Reinforcement Learning14 Feb 2021 0 repositories listed
-
A Reinforcement learning method for Optical Thin-Film Design13 Feb 2021 0 repositories listed
-
Equilibrium Inverse Reinforcement Learning for Ride-hailing Vehicle Network13 Feb 2021 0 repositories listed
-
Improved Corruption Robust Algorithms for Episodic Reinforcement Learning13 Feb 2021 0 repositories listed
-
Modelling Cooperation in Network Games with Spatio-Temporal Complexity13 Feb 2021 0 repositories listed
-
PerSim: Data-Efficient Offline Reinforcement Learning with Heterogeneous Agents via Personalized Simulators13 Feb 2021 0 repositories listed
-
Deep Reinforcement Learning for Backup Strategies against Adversaries12 Feb 2021 0 repositories listed
-
Discovery of Options via Meta-Learned Subgoals12 Feb 2021 0 repositories listed
-
Disturbing Reinforcement Learning Agents with Corrupted Rewards12 Feb 2021 0 repositories listed
-
Reinforcement Learning For Data Poisoning on Graph Neural Networks12 Feb 2021 0 repositories listed
-
Deep Reinforcement Learning for Combinatorial Optimization: Covering Salesman Problems11 Feb 2021 0 repositories listed
-
Deep Reinforcement Learning for Portfolio Optimization using Latent Feature State Space (LFSS) Module11 Feb 2021 0 repositories listed
-
Hedging of Financial Derivative Contracts via Monte Carlo Tree Search11 Feb 2021 0 repositories listed