Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 125
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 125 of 152: papers 12,401 to 12,500 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Deep Reinforcement Learning-Based Beam Tracking for Low-Latency Services in Vehicular Networks13 Feb 2020 0 repositories listed
-
Fast Reinforcement Learning for Anti-jamming Communications13 Feb 2020 0 repositories listed
-
Improving Generalization of Reinforcement Learning with Minimax Distributional Soft Actor-Critic13 Feb 2020 0 repositories listed
-
MODRL/D-AM: Multiobjective Deep Reinforcement Learning Algorithm Using Decomposition and Attention Model for Multiobjective Optimization13 Feb 2020 0 repositories listed
-
Multi-Vehicle Routing Problems with Soft Time Windows: A Multi-Agent Reinforcement Learning Approach13 Feb 2020 0 repositories listed
-
A Tensor Network Approach to Finite Markov Decision Processes12 Feb 2020 0 repositories listed
-
Data Efficient Training for Reinforcement Learning with Adaptive Behavior Policy Sharing12 Feb 2020 0 repositories listed
-
Regret Bounds for Discounted MDPs12 Feb 2020 0 repositories listed
-
Confounding-Robust Policy Evaluation in Infinite-Horizon Reinforcement Learning11 Feb 2020 0 repositories listed
-
HMRL: Hyper-Meta Learning for Sparse Reward Reinforcement Learning Problem11 Feb 2020 0 repositories listed
-
Learning Structured Communication for Multi-agent Reinforcement Learning11 Feb 2020 0 repositories listed
-
Learning to Switch Among Agents in a Team via 2-Layer Markov Decision Processes11 Feb 2020 0 repositories listed
-
Machine Learning Approaches For Motor Learning: A Short Review11 Feb 2020 0 repositories listed
-
Towards Intelligent Pick and Place Assembly of Individualized Products Using Reinforcement Learning11 Feb 2020 0 repositories listed
-
Interpretable Off-Policy Evaluation in Reinforcement Learning by Highlighting Influential Transitions10 Feb 2020 0 repositories listed
-
On Reward Shaping for Mobile Robot Navigation: A Reinforcement Learning and SLAM Based Approach10 Feb 2020 0 repositories listed
-
On the Convergence of the Monte Carlo Exploring Starts Algorithm for Reinforcement Learning10 Feb 2020 0 repositories listed
-
Proficiency Constrained Multi-Agent Reinforcement Learning for Environment-Adaptive Multi UAV-UGV Teaming10 Feb 2020 0 repositories listed
-
Provable Self-Play Algorithms for Competitive Reinforcement Learning10 Feb 2020 0 repositories listed
-
Reward Tweaking: Maximizing the Total Reward While Planning for Short Horizons9 Feb 2020 0 repositories listed
-
A data-driven choice of misfit function for FWI using reinforcement learning8 Feb 2020 0 repositories listed
-
8 Feb 2020 0 repositories listed
-
BRPO: Batch Residual Policy Optimization8 Feb 2020 0 repositories listed
-
Conservative Exploration in Reinforcement Learning8 Feb 2020 0 repositories listed
-
Description Based Text Classification with Reinforcement Learning8 Feb 2020 0 repositories listed
-
Inferential Induction: A Novel Framework for Bayesian Reinforcement Learning8 Feb 2020 0 repositories listed
-
Multi-task Reinforcement Learning with a Planning Quasi-Metric8 Feb 2020 0 repositories listed
-
RL-Duet: Online Music Accompaniment Generation Using Deep Reinforcement Learning8 Feb 2020 0 repositories listed
-
Accelerating Reinforcement Learning for Reaching using Continuous Curriculum Learning7 Feb 2020 0 repositories listed
-
Automated Lane Change Strategy using Proximal Policy Optimization-based Deep Reinforcement Learning7 Feb 2020 0 repositories listed
-
Bayesian Residual Policy Optimization: Scalable Bayesian Reinforcement Learning with Clairvoyant Experts7 Feb 2020 0 repositories listed
-
Causally Correct Partial Models for Reinforcement Learning7 Feb 2020 0 repositories listed
-
Explicit Mean-Square Error Bounds for Monte-Carlo and Linear Stochastic Approximation7 Feb 2020 0 repositories listed
-
Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals7 Feb 2020 0 repositories listed
-
Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces7 Feb 2020 0 repositories listed
-
Reward-Free Exploration for Reinforcement Learning7 Feb 2020 0 repositories listed
-
Student/Teacher Advising through Reward Augmentation7 Feb 2020 0 repositories listed
-
Reinforcement Learning in Factored MDPs: Oracle-Efficient Algorithms and Tighter Regret Bounds for the Non-Episodic Setting6 Feb 2020 0 repositories listed
-
Social diversity and social preferences in mixed-motive reinforcement learning6 Feb 2020 0 repositories listed
-
Temporal-adaptive Hierarchical Reinforcement Learning6 Feb 2020 0 repositories listed
-
Deep Radial-Basis Value Functions for Continuous Control5 Feb 2020 0 repositories listed
-
Mutual Information-based State-Control for Intrinsically Motivated Reinforcement Learning5 Feb 2020 0 repositories listed
-
Bootstrapping a DQN Replay Memory with Synthetic Experiences4 Feb 2020 0 repositories listed
-
Finite Time Analysis of Linear Two-timescale Stochastic Approximation with Markovian Noise4 Feb 2020 0 repositories listed
-
Learning Task-Driven Control Policies via Information Bottlenecks4 Feb 2020 0 repositories listed
-
Policy Gradient based Quantum Approximate Optimization Algorithm4 Feb 2020 0 repositories listed
-
Evolutionary algorithms for constructing an ensemble of decision trees3 Feb 2020 0 repositories listed
-
Finite-Sample Analysis of Stochastic Approximation Using Smooth Convex Envelopes3 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning for Autonomous Driving: A Survey2 Feb 2020 0 repositories listed
-
PolicyGNN: Aggregation Optimization for Graph Neural Networks1 Feb 2020 0 repositories listed
-
A Deep Reinforcement Learning Approach to Concurrent Bilateral Negotiation31 Jan 2020 0 repositories listed
-
Locally Private Distributed Reinforcement Learning31 Jan 2020 0 repositories listed
-
Predicting Goal-directed Attention Control Using Inverse-Reinforcement Learning31 Jan 2020 0 repositories listed
-
Preventing Imitation Learning with Adversarial Policy Ensembles31 Jan 2020 0 repositories listed
-
Survey of Deep Reinforcement Learning for Motion Planning of Autonomous Vehicles30 Jan 2020 0 repositories listed
-
Asymptotically Efficient Off-Policy Evaluation for Tabular Reinforcement Learning29 Jan 2020 0 repositories listed
-
Robust Multimodal Image Registration Using Deep Recurrent Reinforcement Learning29 Jan 2020 0 repositories listed
-
Data-driven control of micro-climate in buildings: an event-triggered reinforcement learning approach28 Jan 2020 0 repositories listed
-
Distal Explanations for Model-free Explainable Reinforcement Learning28 Jan 2020 0 repositories listed
-
Developing Multi-Task Recommendations with Long-Term Rewards via Policy Distilled Reinforcement Learning27 Jan 2020 0 repositories listed
-
Reinforcement Learning-based Application Autoscaling in the Cloud: A Survey27 Jan 2020 0 repositories listed
-
Unsupervised Program Synthesis for Images By Sampling Without Replacement27 Jan 2020 0 repositories listed
-
Constrained Upper Confidence Reinforcement Learning26 Jan 2020 0 repositories listed
-
Sentiment and Knowledge Based Algorithmic Trading with Deep Reinforcement Learning26 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning based Blind mmWave MIMO Beam Alignment25 Jan 2020 0 repositories listed
-
Following Instructions by Imagining and Reaching Visual Goals25 Jan 2020 0 repositories listed
-
EgoMap: Projective mapping and structured egocentric memory for Deep RL24 Jan 2020 0 repositories listed
-
End-to-End Vision-Based Adaptive Cruise Control (ACC) Using Deep Reinforcement Learning24 Jan 2020 0 repositories listed
-
Pricing commodity swing options24 Jan 2020 0 repositories listed
-
Facial Feedback for Reinforcement Learning: A Case Study and Offline Analysis Using the TAMER Framework23 Jan 2020 0 repositories listed
-
Reducing Non-Normative Text Generation from Language Models23 Jan 2020 0 repositories listed
-
Multi-objective Neural Architecture Search via Non-stationary Policy Gradient23 Jan 2020 0 repositories listed
-
Local Policy Optimization for Trajectory-Centric Reinforcement Learning22 Jan 2020 0 repositories listed
-
On Solving Cooperative MARL Problems with a Few Good Experiences22 Jan 2020 0 repositories listed
-
Reinforcement Learning Based Vehicle-cell Association Algorithm for Highly Mobile Millimeter Wave Communication22 Jan 2020 0 repositories listed
-
Cooperative Highway Work Zone Merge Control based on Reinforcement Learning in A Connected and Automated Environment21 Jan 2020 0 repositories listed
-
Improving Interaction Quality Estimation with BiLSTMs and the Impact on Dialogue Policy Learning21 Jan 2020 0 repositories listed
-
Intelligent Bandwidth Allocation for Latency Management in NG-EPON using Reinforcement Learning Methods21 Jan 2020 0 repositories listed
-
Lyceum: An efficient and scalable ecosystem for robot learning21 Jan 2020 0 repositories listed
-
Unsupervisedly Learned Representations: Should the Quest be Over?21 Jan 2020 0 repositories listed
-
Memristor Hardware-Friendly Reinforcement Learning20 Jan 2020 0 repositories listed
-
Nested-Wasserstein Self-Imitation Learning for Sequence Generation20 Jan 2020 0 repositories listed
-
Reinforcement Learning with Probabilistically Complete Exploration20 Jan 2020 0 repositories listed
-
A Survey of Reinforcement Learning Techniques: Strategies, Recent Development, and Future Directions19 Jan 2020 0 repositories listed
-
FRESH: Interactive Reward Shaping in High-Dimensional State Spaces using Human Feedback19 Jan 2020 0 repositories listed
-
Learning Options from Demonstration using Skill Segmentation19 Jan 2020 0 repositories listed
-
cube2net: Efficient Query-Specific Network Construction with Data Cube Organization18 Jan 2020 0 repositories listed
-
BNAS:An Efficient Neural Architecture Search Approach Using Broad Scalable Architecture18 Jan 2020 0 repositories listed
-
Multi-agent Motion Planning for Dense and Dynamic Environments via Deep Reinforcement Learning18 Jan 2020 0 repositories listed
-
Algorithms in Multi-Agent Systems: A Holistic Perspective from Reinforcement Learning and Game Theory17 Jan 2020 0 repositories listed
-
MIME: Mutual Information Minimisation Exploration16 Jan 2020 0 repositories listed
-
Reward Shaping for Reinforcement Learning with Omega-Regular Objectives16 Jan 2020 0 repositories listed
-
Model-based Multi-Agent Reinforcement Learning with Cooperative Prioritized Sweeping15 Jan 2020 0 repositories listed
-
Robotic Grasp Manipulation Using Evolutionary Computing and Deep Reinforcement Learning15 Jan 2020 0 repositories listed
-
SEERL: Sample Efficient Ensemble Reinforcement Learning15 Jan 2020 0 repositories listed
-
Exploiting Language Instructions for Interpretable and Compositional Reinforcement Learning13 Jan 2020 0 repositories listed
-
Learning to Locomote with Deep Neural-Network and CPG-based Control in a Soft Snake Robot13 Jan 2020 0 repositories listed
-
Multi-Robot Formation Control Using Reinforcement Learning13 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning for Complex Manipulation Tasks with Sparse Feedback12 Jan 2020 0 repositories listed
-
Weakly Supervised Video Summarization by Hierarchical Reinforcement Learning12 Jan 2020 0 repositories listed