Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 84
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 84 of 152: papers 8,301 to 8,400 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Empathetic Persuasion: Reinforcing Empathy and Persuasiveness in Dialogue Systems1 Jul 2022 0 repositories listed
-
Interactive Learning from Natural Language and Demonstrations using Signal Temporal Logic1 Jul 2022 0 repositories listed
-
Partner Personas Generation for Dialogue Response Generation1 Jul 2022 0 repositories listed
-
Reinforcement Learning Based User-Guided Motion Planning for Human-Robot Collaboration1 Jul 2022 0 repositories listed
-
Reinforcement Learning of Multi-Domain Dialog Policies Via Action Embeddings1 Jul 2022 0 repositories listed
-
Safe Decision-making for Lane-change of Autonomous Vehicles via Human Demonstration-aided Reinforcement Learning1 Jul 2022 0 repositories listed
-
SURF: Semantic-level Unsupervised Reward Function for Machine Translation1 Jul 2022 0 repositories listed
-
Depth-CUPRL: Depth-Imaged Contrastive Unsupervised Prioritized Representations in Reinforcement Learning for Mapless Navigation of Unmanned Aerial Vehicles30 Jun 2022 0 repositories listed
-
Performative Reinforcement Learning30 Jun 2022 0 repositories listed
-
Deep Reinforcement Learning for Small Bowel Path Tracking using Different Types of Annotations29 Jun 2022 0 repositories listed
-
Minimalist and High-performance Conversational Recommendation with Uncertainty Estimation for User Preference29 Jun 2022 0 repositories listed
-
Provably Efficient Reinforcement Learning for Online Adaptive Influence Maximization29 Jun 2022 0 repositories listed
-
Applications of Reinforcement Learning in Finance -- Trading with a Double Deep Q-Network28 Jun 2022 0 repositories listed
-
Dependency Parsing with Backtracking using Deep Reinforcement Learning28 Jun 2022 0 repositories listed
-
GAN-based Intrinsic Exploration For Sample Efficient Reinforcement Learning28 Jun 2022 0 repositories listed
-
Masked World Models for Visual Control28 Jun 2022 0 repositories listed
-
Position-Agnostic Autonomous Navigation in Vineyards with Deep Reinforcement Learning28 Jun 2022 0 repositories listed
-
Reinforcement Learning Based Dynamic Model Combination for Time Series Forecasting28 Jun 2022 0 repositories listed
-
Reinforcement Learning in Medical Image Analysis: Concepts, Applications, Challenges, and Future Directions28 Jun 2022 0 repositories listed
-
Risk Perspective Exploration in Distributional Reinforcement Learning28 Jun 2022 0 repositories listed
-
Spatial Positioning Token (SPToken) for Smart Parking28 Jun 2022 0 repositories listed
-
Traffic Management of Autonomous Vehicles using Policy Based Deep Reinforcement Learning and Intelligent Routing28 Jun 2022 0 repositories listed
-
EMVLight: a Multi-agent Reinforcement Learning Framework for an Emergency Vehicle Decentralized Routing and Traffic Signal Control System27 Jun 2022 0 repositories listed
-
Humans are not Boltzmann Distributions: Challenges and Opportunities for Modelling Human Feedback and Interaction in Reinforcement Learning27 Jun 2022 0 repositories listed
-
Interpretable Hidden Markov Model-Based Deep Reinforcement Learning Hierarchical Framework for Predictive Maintenance of Turbofan Engines27 Jun 2022 0 repositories listed
-
On the Complexity of Adversarial Decision Making27 Jun 2022 0 repositories listed
-
Analysis of Stochastic Processes through Replay Buffers26 Jun 2022 0 repositories listed
-
Estimating Link Flows in Road Networks with Synthetic Trajectory Data Generation: Reinforcement Learning-based Approaches26 Jun 2022 0 repositories listed
-
Predicting the Need for Blood Transfusion in Intensive Care Units with Reinforcement Learning26 Jun 2022 0 repositories listed
-
Functional Optimization Reinforcement Learning for Real-Time Bidding25 Jun 2022 0 repositories listed
-
Hierarchical Reinforcement Learning with Opponent Modeling for Distributed Multi-agent Cooperation25 Jun 2022 0 repositories listed
-
Towards Modern Card Games with Large-Scale Action Spaces Through Action Representation25 Jun 2022 0 repositories listed
-
Value-Consistent Representation Learning for Data-Efficient Reinforcement Learning25 Jun 2022 0 repositories listed
-
Dynamic network congestion pricing based on deep reinforcement learning24 Jun 2022 0 repositories listed
-
Eco-driving for Electric Connected Vehicles at Signalized Intersections: A Parameterized Reinforcement Learning approach24 Jun 2022 0 repositories listed
-
Joint Representation Training in Sequential Tasks with Shared Structure24 Jun 2022 0 repositories listed
-
Learning the policy for mixed electric platoon control of automated and human-driven vehicles at signalized intersection: a random search approach24 Jun 2022 0 repositories listed
-
Modeling Adaptive Platoon and Reservation Based Autonomous Intersection Control: A Deep Reinforcement Learning Approach24 Jun 2022 0 repositories listed
-
Phasic Self-Imitative Reduction for Sparse-Reward Goal-Conditioned Reinforcement Learning24 Jun 2022 0 repositories listed
-
Provably Efficient Reinforcement Learning in Partially Observable Dynamical Systems24 Jun 2022 0 repositories listed
-
Value Function Decomposition for Iterative Design of Reinforcement Learning Agents24 Jun 2022 0 repositories listed
-
A Federated Reinforcement Learning Method with Quantization for Cooperative Edge Caching in Fog Radio Access Networks23 Jun 2022 0 repositories listed
-
Learning Agile Skills via Adversarial Imitation of Rough Partial Demonstrations23 Jun 2022 0 repositories listed
-
Nearly Minimax Optimal Reinforcement Learning with Linear Function Approximation23 Jun 2022 0 repositories listed
-
Recursive Reinforcement Learning23 Jun 2022 0 repositories listed
-
Reinforcement Learning under Partial Observability Guided by Learned Environment Models23 Jun 2022 0 repositories listed
-
The Real Deal: A Review of Challenges and Opportunities in Moving Reinforcement Learning-Based Traffic Signal Control Systems Towards Reality23 Jun 2022 0 repositories listed
-
Auto-Encoding Adversarial Imitation Learning22 Jun 2022 0 repositories listed
-
Curious Exploration via Structured World Models Yields Zero-Shot Object Manipulation22 Jun 2022 0 repositories listed
-
Decentralized Gossip-Based Stochastic Bilevel Optimization over Communication Networks22 Jun 2022 0 repositories listed
-
Fusion of Model-free Reinforcement Learning with Microgrid Control: Review and Vision22 Jun 2022 0 repositories listed
-
Learning Optimal Treatment Strategies for Sepsis Using Offline Reinforcement Learning in Continuous Space22 Jun 2022 0 repositories listed
-
Constrained Stochastic Nonconvex Optimization with State-dependent Markov Data22 Jun 2022 0 repositories listed
-
A Single-Timescale Analysis For Stochastic Approximation With Multiple Coupled Sequences21 Jun 2022 0 repositories listed
-
Federated Stochastic Approximation under Markov Noise and Heterogeneity: Applications in Reinforcement Learning21 Jun 2022 0 repositories listed
-
Finding Optimal Policy for Queueing Models: New Parameterization21 Jun 2022 0 repositories listed
-
Hybridization of evolutionary algorithm and deep reinforcement learning for multi-objective orienteering optimization21 Jun 2022 0 repositories listed
-
Imitate then Transcend: Multi-Agent Optimal Execution with Dual-Window Denoise PPO21 Jun 2022 0 repositories listed
-
Incorporating Voice Instructions in Model-Based Reinforcement Learning for Self-Driving Cars21 Jun 2022 0 repositories listed
-
Model-Based Imitation Learning Using Entropy Regularization of Model and Policy21 Jun 2022 0 repositories listed
-
On the Statistical Efficiency of Reward-Free Exploration in Non-Linear RL21 Jun 2022 0 repositories listed
-
Safe and Psychologically Pleasant Traffic Signal Control with Reinforcement Learning using Action Masking21 Jun 2022 0 repositories listed
-
The Integration of Machine Learning into Automated Test Generation: A Systematic Mapping Study21 Jun 2022 0 repositories listed
-
Constrained Reinforcement Learning for Robotics via Scenario-Based Programming20 Jun 2022 0 repositories listed
-
Deep reinforced active learning for multi-class image classification20 Jun 2022 0 repositories listed
-
From Multi-agent to Multi-robot: A Scalable Training and Evaluation Platform for Multi-robot Reinforcement Learning20 Jun 2022 0 repositories listed
-
Guided Safe Shooting: model based reinforcement learning with safety constraints20 Jun 2022 0 repositories listed
-
S2RL: Do We Really Need to Perceive All States in Deep Multi-Agent Reinforcement Learning?20 Jun 2022 0 repositories listed
-
A Survey on Model-based Reinforcement Learning19 Jun 2022 0 repositories listed
-
Guarantees for Epsilon-Greedy Reinforcement Learning with Function Approximation19 Jun 2022 0 repositories listed
-
Learning Multi-Task Transferable Rewards via Variational Inverse Reinforcement Learning19 Jun 2022 0 repositories listed
-
Two-Hop Age of Information Scheduling for Multi-UAV Assisted Mobile Edge Computing: FRL vs MADDPG19 Jun 2022 0 repositories listed
-
AnyMorph: Learning Transferable Polices By Inferring Agent Morphology17 Jun 2022 0 repositories listed
-
Deep reinforcement learning for fMRI prediction of Autism Spectrum Disorder17 Jun 2022 0 repositories listed
-
Generalised Policy Improvement with Geometric Policy Composition17 Jun 2022 0 repositories listed
-
Backbones-Review: Feature Extraction Networks for Deep Learning and Deep Reinforcement Learning Approaches16 Jun 2022 0 repositories listed
-
Reinforcement Learning for Economic Policy: A New Frontier?16 Jun 2022 0 repositories listed
-
A Look at Value-Based Decision-Time vs. Background Planning Methods Across Different Settings16 Jun 2022 0 repositories listed
-
Automating the resolution of flight conflicts: Deep reinforcement learning in service of air traffic controllers15 Jun 2022 0 repositories listed
-
Autonomous Platoon Control with Integrated Deep Reinforcement Learning and Dynamic Programming15 Jun 2022 0 repositories listed
-
Contrastive Learning as Goal-Conditioned Reinforcement Learning15 Jun 2022 0 repositories listed
-
Mean-Semivariance Policy Optimization via Risk-Averse Reinforcement Learning15 Jun 2022 0 repositories listed
-
Rethinking Reinforcement Learning for Recommendation: A Prompt Perspective15 Jun 2022 0 repositories listed
-
Revisiting Some Common Practices in Cooperative Multi-Agent Reinforcement Learning15 Jun 2022 0 repositories listed
-
Deep Reinforcement Learning for Exact Combinatorial Optimization: Learning to Branch14 Jun 2022 0 repositories listed
-
FreeKD: Free-direction Knowledge Distillation for Graph Neural Networks14 Jun 2022 0 repositories listed
-
Open-Ended Learning Strategies for Learning Complex Locomotion Skills14 Jun 2022 0 repositories listed
-
Robust Reinforcement Learning with Distributional Risk-averse formulation14 Jun 2022 0 repositories listed
-
Solving the capacitated vehicle routing problem with timing windows using rollouts and MAX-SAT14 Jun 2022 0 repositories listed
-
Stein Variational Goal Generation for adaptive Exploration in Multi-Goal Reinforcement Learning14 Jun 2022 0 repositories listed
-
Towards a Solution to Bongard Problems: A Causal Approach14 Jun 2022 0 repositories listed
-
Variance Reduction for Policy-Gradient Methods via Empirical Variance Minimization14 Jun 2022 0 repositories listed
-
Visual Radial Basis Q-Network14 Jun 2022 0 repositories listed
-
Analysis of Randomization Effects on Sim2Real Transfer in Reinforcement Learning for Robotic Manipulation Tasks13 Jun 2022 0 repositories listed
-
Computation Offloading and Resource Allocation in F-RANs: A Federated Deep Reinforcement Learning Approach13 Jun 2022 0 repositories listed
-
Intrinsically motivated option learning: a comparative study of recent methods13 Jun 2022 0 repositories listed
-
Provable Benefit of Multitask Representation Learning in Reinforcement Learning13 Jun 2022 0 repositories listed
-
Provably Efficient Offline Reinforcement Learning with Trajectory-Wise Reward13 Jun 2022 0 repositories listed
-
Relative Policy-Transition Optimization for Fast Policy Transfer13 Jun 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.