Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 96
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 96 of 152: papers 9,501 to 9,600 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Self-Consistent Models and Values25 Oct 2021 0 repositories listed
-
Unsupervised Domain Adaptation with Dynamics-Aware Rewards in Reinforcement Learning25 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Simultaneous Sensing and Channel Access in Cognitive Networks24 Oct 2021 0 repositories listed
-
Analysis of Thompson Sampling for Partially Observable Contextual Multi-Armed Bandits23 Oct 2021 0 repositories listed
-
Foresight of Graph Reinforcement Learning Latent Permutations Learnt by Gumbel Sinkhorn Network23 Oct 2021 0 repositories listed
-
Fully Distributed Actor-Critic Architecture for Multitask Deep Reinforcement Learning23 Oct 2021 0 repositories listed
-
Policy Search using Dynamic Mirror Descent MPC for Model Free Off Policy RL23 Oct 2021 0 repositories listed
-
A Reinforcement Learning Approach to Parameter Selection for Distributed Optimal Power Flow22 Oct 2021 0 repositories listed
-
C-Planning: An Automatic Curriculum for Learning Goal-Reaching Tasks22 Oct 2021 0 repositories listed
-
Convergence Rates of Average-Reward Multi-agent Reinforcement Learning via Randomized Linear Programming22 Oct 2021 0 repositories listed
-
Off-policy Reinforcement Learning with Optimistic Exploration and Distribution Correction22 Oct 2021 0 repositories listed
-
Patient level simulation and reinforcement learning to discover novel strategies for treating ovarian cancer22 Oct 2021 0 repositories listed
-
Reinforcement Learning for Process Control with Application in Semiconductor Manufacturing22 Oct 2021 0 repositories listed
-
Anti-Concentrated Confidence Bonuses for Scalable Exploration21 Oct 2021 0 repositories listed
-
Can Q-learning solve Multi Armed Bantids?21 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Online Control of Stochastic Partial Differential Equations21 Oct 2021 0 repositories listed
-
Efficient Robotic Manipulation Through Offline-to-Online Reinforcement Learning and Goal-Aware State Information21 Oct 2021 0 repositories listed
-
Is High Variance Unavoidable in RL? A Case Study in Continuous Control21 Oct 2021 0 repositories listed
-
Model-based Reinforcement Learning for Service Mesh Fault Resiliency in a Web Application-level21 Oct 2021 0 repositories listed
-
Neuro-Symbolic Reinforcement Learning with First-Order Logic21 Oct 2021 0 repositories listed
-
Off-Dynamics Inverse Reinforcement Learning from Hetero-Domain21 Oct 2021 0 repositories listed
-
Reinforcement Learning Based Optimal Camera Placement for Depth Observation of Indoor Scenes21 Oct 2021 0 repositories listed
-
Socialbots on Fire: Modeling Adversarial Behaviors of Socialbots via Multi-Agent Hierarchical Reinforcement Learning20 Oct 2021 0 repositories listed
-
Computationally Efficient Safe Reinforcement Learning for Power Systems20 Oct 2021 0 repositories listed
-
Distributed Reinforcement Learning for Privacy-Preserving Dynamic Edge Caching20 Oct 2021 0 repositories listed
-
Feedback Linearization of Car Dynamics for Racing via Reinforcement Learning20 Oct 2021 0 repositories listed
-
More Efficient Exploration with Symbolic Priors on Action Sequence Equivalences20 Oct 2021 0 repositories listed
-
Transferring Reinforcement Learning for DC-DC Buck Converter Control via Duty Ratio Mapping: From Simulation to Implementation20 Oct 2021 0 repositories listed
-
Aesthetic Photo Collage with Deep Reinforcement Learning19 Oct 2021 0 repositories listed
-
Beyond Exact Gradients: Convergence of Stochastic Soft-Max Policy Gradient Methods with Entropy Regularization19 Oct 2021 0 repositories listed
-
Improved cooperation by balancing exploration and exploitation in intertemporal social dilemma tasks19 Oct 2021 0 repositories listed
-
Learning Robotic Manipulation Skills Using an Adaptive Force-Impedance Action Space19 Oct 2021 0 repositories listed
-
Locally Differentially Private Reinforcement Learning for Linear Mixture Markov Decision Processes19 Oct 2021 0 repositories listed
-
Neural Network Compatible Off-Policy Natural Actor-Critic Algorithm19 Oct 2021 0 repositories listed
-
On Reward-Free RL with Kernel and Neural Function Approximations: Single-Agent MDP and Markov Game19 Oct 2021 0 repositories listed
-
State-based Episodic Memory for Multi-Agent Reinforcement Learning19 Oct 2021 0 repositories listed
-
Embracing advanced AI/ML to help investors achieve success: Vanguard Reinforcement Learning for Financial Goal Planning18 Oct 2021 0 repositories listed
-
Improving Robustness of Reinforcement Learning for Power System Control with Adversarial Training18 Oct 2021 0 repositories listed
-
Option Transfer and SMDP Abstraction with Successor Features18 Oct 2021 0 repositories listed
-
Optimistic Policy Optimization is Provably Efficient in Non-stationary MDPs18 Oct 2021 0 repositories listed
-
Provable Hierarchy-Based Meta-Reinforcement Learning18 Oct 2021 0 repositories listed
-
Reinforcement Learning-Based Coverage Path Planning with Implicit Cellular Decomposition18 Oct 2021 0 repositories listed
-
Sim-to-Real Transfer in Multi-agent Reinforcement Networking for Federated Edge Computing18 Oct 2021 0 repositories listed
-
Damped Anderson Mixing for Deep Reinforcement Learning: Acceleration, Convergence, and Stabilization17 Oct 2021 0 repositories listed
-
Provable RL with Exogenous Distractors via Multistep Inverse Dynamics17 Oct 2021 0 repositories listed
-
Towards Instance-Optimal Offline Reinforcement Learning with Pessimism17 Oct 2021 0 repositories listed
-
Case-based Reasoning for Better Generalization in Textual Reinforcement Learning16 Oct 2021 0 repositories listed
-
Emotion Style Transfer with a Specified Intensity Using Deep Reinforcement Learning16 Oct 2021 0 repositories listed
-
Generative Adversarial Imitation Learning for End-to-End Autonomous Driving on Urban Environments16 Oct 2021 0 repositories listed
-
Hallucinated but Factual! Inspecting the Factuality of Hallucinations in Abstractive Summarization16 Oct 2021 0 repositories listed
-
Lifting the veil on hyper-parameters for value-based deep reinforcement learning16 Oct 2021 0 repositories listed
-
Local Advantage Actor-Critic for Robust Multi-Agent Deep Reinforcement Learning16 Oct 2021 0 repositories listed
-
Neural Network Pruning Through Constrained Reinforcement Learning16 Oct 2021 0 repositories listed
-
Online Target Q-learning with Reverse Experience Replay: Efficiently finding the Optimal Policy for Linear MDPs16 Oct 2021 0 repositories listed
-
Rethinking Modern Communication from Semantic Coding to Semantic Communication16 Oct 2021 0 repositories listed
-
A Broad-persistent Advising Approach for Deep Interactive Reinforcement Learning in Robotic Environments15 Oct 2021 0 repositories listed
-
Containerized Distributed Value-Based Multi-Agent Reinforcement Learning15 Oct 2021 0 repositories listed
-
Dynamic probabilistic logic models for effective abstractions in RL15 Oct 2021 0 repositories listed
-
Effects of Different Optimization Formulations in Evolutionary Reinforcement Learning on Diverse Behavior Generation15 Oct 2021 0 repositories listed
-
GrowSpace: Learning How to Shape Plants15 Oct 2021 0 repositories listed
-
Improving Hyperparameter Optimization by Planning Ahead15 Oct 2021 0 repositories listed
-
On-Policy Model Errors in Reinforcement Learning15 Oct 2021 0 repositories listed
-
Value Penalized Q-Learning for Recommender Systems15 Oct 2021 0 repositories listed
-
Wasserstein Unsupervised Reinforcement Learning15 Oct 2021 0 repositories listed
-
HAVEN: Hierarchical Cooperative Multi-Agent Reinforcement Learning with Dual Coordination Mechanism14 Oct 2021 0 repositories listed
-
Offline Reinforcement Learning with Soft Behavior Regularization14 Oct 2021 0 repositories listed
-
Provably Efficient Multi-Agent Reinforcement Learning with Fully Decentralized Communication14 Oct 2021 0 repositories listed
-
Safe Autonomous Racing via Approximate Reachability on Ego-vision14 Oct 2021 0 repositories listed
-
Sign and Relevance Learning14 Oct 2021 0 repositories listed
-
Block Contextual MDPs for Continual Learning13 Oct 2021 0 repositories listed
-
Feudal Reinforcement Learning by Reading Manuals13 Oct 2021 0 repositories listed
-
NeurIPS 2021 Competition IGLU: Interactive Grounded Language Understanding in a Collaborative Environment13 Oct 2021 0 repositories listed
-
On Covariate Shift of Latent Confounders in Imitation and Reinforcement Learning13 Oct 2021 0 repositories listed
-
Reinforcement Learning for Standards Design13 Oct 2021 0 repositories listed
-
On Improving Model-Free Algorithms for Decentralized Multi-Agent Reinforcement Learning12 Oct 2021 0 repositories listed
-
Deciding What's Fair: Challenges of Applying Reinforcement Learning in Online Marketplaces12 Oct 2021 0 repositories listed
-
DQN-based Beamforming for Uplink mmWave Cellular-Connected UAVs12 Oct 2021 0 repositories listed
-
GridLearn: Multiagent Reinforcement Learning for Grid-Aware Building Energy Management12 Oct 2021 0 repositories listed
-
Learning Efficient Multi-Agent Cooperative Visual Exploration12 Oct 2021 0 repositories listed
-
Power and Accountability in RL-driven Environmental Policy12 Oct 2021 0 repositories listed
-
Provably Efficient Reinforcement Learning in Decentralized General-Sum Markov Games12 Oct 2021 0 repositories listed
-
Reward-Free Model-Based Reinforcement Learning with Linear Function Approximation12 Oct 2021 0 repositories listed
-
Temporal Abstraction in Reinforcement Learning with the Successor Representation12 Oct 2021 0 repositories listed
-
An Automated Portfolio Trading System with Feature Preprocessing and Recurrent Reinforcement Learning11 Oct 2021 0 repositories listed
-
Bid Optimization using Maximum Entropy Reinforcement Learning11 Oct 2021 0 repositories listed
-
Navigation In Urban Environments Amongst Pedestrians Using Multi-Objective Deep Reinforcement Learning11 Oct 2021 0 repositories listed
-
REIN-2: Giving Birth to Prepared Reinforcement Learning Agents Using Reinforcement Learning Agents11 Oct 2021 0 repositories listed
-
Scalable Traffic Signal Controls using Fog-Cloud Based Multiagent Reinforcement Learning11 Oct 2021 0 repositories listed
-
An Augmented Reality Platform for Introducing Reinforcement Learning to K-12 Students with Robots10 Oct 2021 0 repositories listed
-
An In-depth Summary of Recent Artificial Intelligence Applications in Drug Design10 Oct 2021 0 repositories listed
-
Braxlines: Fast and Interactive Toolkit for RL-driven Behavior Engineering beyond Reward Maximization10 Oct 2021 0 repositories listed
-
Deep Reinforcement Learning for Optimizing RIS-Assisted HD-FD Wireless Systems10 Oct 2021 0 repositories listed
-
Hard instance learning for quantum adiabatic prime factorization10 Oct 2021 0 repositories listed
-
Multi-condition multi-objective optimization using deep reinforcement learning10 Oct 2021 0 repositories listed
-
Reinforcement Learning for Systematic FX Trading10 Oct 2021 0 repositories listed
-
Satisficing Paths and Independent Multi-Agent Reinforcement Learning in Stochastic Games9 Oct 2021 0 repositories listed
-
Breaking the Sample Complexity Barrier to Regret-Optimal Model-Free Reinforcement Learning9 Oct 2021 0 repositories listed
-
Human-Aware Robot Navigation via Reinforcement Learning with Hindsight Experience Replay and Curriculum Learning9 Oct 2021 0 repositories listed
-
Provably Efficient Black-Box Action Poisoning Attacks Against Reinforcement Learning9 Oct 2021 0 repositories listed
-
Theoretically Principled Deep RL Acceleration via Nearest Neighbor Function Approximation9 Oct 2021 0 repositories listed