Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 92
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 92 of 152: papers 9,101 to 9,200 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
State of the Art of Reinforcement Learning17 Jan 2022 0 repositories listed
-
Summarising and Comparing Agent Dynamics with Contrastive Spatiotemporal Abstraction17 Jan 2022 0 repositories listed
-
Towards deep observation: A systematic survey on artificial intelligence techniques to monitor fetus via Ultrasound Images17 Jan 2022 0 repositories listed
-
A Family of Cognitively Realistic Parsing Environments for Deep Reinforcement Learning16 Jan 2022 0 repositories listed
-
CONQRR: Conversational Query Rewriting for Retrieval with Reinforcement Learning16 Jan 2022 0 repositories listed
-
Inherently Explainable Reinforcement Learning in Natural Language16 Jan 2022 0 repositories listed
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture16 Jan 2022 0 repositories listed
-
Learning from Atypical Behavior: Temporary Interest Aware Recommendation Based on Reinforcement Learning16 Jan 2022 0 repositories listed
-
Long-Tail Classification for Distinctive Image Captioning: A Simple yet Effective Remedy for Side Effects of Reinforcement Learning16 Jan 2022 0 repositories listed
-
MUST: A Framework for Training Task-oriented Dialogue Systems with Multiple User SimulaTors16 Jan 2022 0 repositories listed
-
Parsing Natural Language into Propositional and First-Order Logic with Dual Reinforcement Learning16 Jan 2022 0 repositories listed
-
Reinforcement Learning with Large Action Spaces for Neural Machine Translation16 Jan 2022 0 repositories listed
-
Revisiting the Roles of “Text” in Text Games16 Jan 2022 0 repositories listed
-
Unsupervised Reinforcement Adaptation for Class-Imbalanced TextClassification16 Jan 2022 0 repositories listed
-
Block Policy Mirror Descent15 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning for Shared Autonomous Vehicles (SAV) Fleet Management15 Jan 2022 0 repositories listed
-
Interpretable and Effective Reinforcement Learning for Attacking against Graph-based Rumor Detection15 Jan 2022 0 repositories listed
-
Profitable Strategy Design by Using Deep Reinforcement Learning for Trades on Cryptocurrency Markets15 Jan 2022 0 repositories listed
-
Recursive Least Squares Advantage Actor-Critic Algorithms15 Jan 2022 0 repositories listed
-
Reinforcement Learning based Air Combat Maneuver Generation14 Jan 2022 0 repositories listed
-
Demystifying Reinforcement Learning in Time-Varying Systems14 Jan 2022 0 repositories listed
-
Reinforcement Learning to Solve NP-hard Problems: an Application to the CVRP14 Jan 2022 0 repositories listed
-
Automated Reinforcement Learning: An Overview13 Jan 2022 0 repositories listed
-
Criticality-Based Varying Step-Number Algorithm for Reinforcement Learning13 Jan 2022 0 repositories listed
-
Direct Mutation and Crossover in Genetic Algorithms Applied to Reinforcement Learning Tasks13 Jan 2022 0 repositories listed
-
Dyna-T: Dyna-Q and Upper Confidence Bounds Applied to Trees12 Jan 2022 0 repositories listed
-
Multi-echelon Supply Chains with Uncertain Seasonal Demands and Lead Times Using Deep Reinforcement Learning12 Jan 2022 0 repositories listed
-
The Recurrent Reinforcement Learning Crypto Agent12 Jan 2022 0 repositories listed
-
Toddler-Guidance Learning: Impacts of Critical Period on Multimodal AI Agents12 Jan 2022 0 repositories listed
-
Active Reinforcement Learning -- A Roadmap Towards Curious Classifier Systems for Self-Adaptation11 Jan 2022 0 repositories listed
-
Automated Reinforcement Learning (AutoRL): A Survey and Open Problems11 Jan 2022 0 repositories listed
-
Benchmarking Deep Reinforcement Learning Algorithms for Vision-based Robotics11 Jan 2022 0 repositories listed
-
Pavlovian Signalling with General Value Functions in Agent-Agent Temporal Decision Making11 Jan 2022 0 repositories listed
-
STIR²: Reward Relabelling for combined Reinforcement and Imitation Learning on sparse-reward tasks11 Jan 2022 0 repositories listed
-
Task Independent Capsule-Based Agents for Deep Q-Learning11 Jan 2022 0 repositories listed
-
Distributed Cooperative Multi-Agent Reinforcement Learning with Directed Coordination Graph10 Jan 2022 0 repositories listed
-
Opportunities of Hybrid Model-based Reinforcement Learning for Cell Therapy Manufacturing Process Control10 Jan 2022 0 repositories listed
-
State of the Art of User Simulation approaches for conversational information retrieval10 Jan 2022 0 repositories listed
-
When is Offline Two-Player Zero-Sum Markov Game Solvable?10 Jan 2022 0 repositories listed
-
A Multi-agent Reinforcement Learning Approach for Efficient Client Selection in Federated Learning9 Jan 2022 0 repositories listed
-
Assessing Policy, Loss and Planning Combinations in Reinforcement Learning using a New Modular Architecture8 Jan 2022 0 repositories listed
-
Neural Network Optimization for Reinforcement Learning Tasks Using Sparse Computations7 Jan 2022 0 repositories listed
-
Offline Reinforcement Learning for Road Traffic Control7 Jan 2022 0 repositories listed
-
Combining Reinforcement Learning and Inverse Reinforcement Learning for Asset Allocation Recommendations6 Jan 2022 0 repositories listed
-
Offsetting Unequal Competition through RL-assisted Incentive Schemes5 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning, a textbook4 Jan 2022 0 repositories listed
-
Learning Complex Spatial Behaviours in ABM: An Experimental Observational Study4 Jan 2022 0 repositories listed
-
A Deeper Understanding of State-Based Critics in Multi-Agent Reinforcement Learning3 Jan 2022 0 repositories listed
-
Actor-Critic Network for Q&A in an Adversarial Environment3 Jan 2022 0 repositories listed
-
Execute Order 66: Targeted Data Poisoning for Reinforcement Learning3 Jan 2022 0 repositories listed
-
Analyzing Micro-Founded General Equilibrium Models with Many Agents using Deep Reinforcement Learning3 Jan 2022 0 repositories listed
-
Reinforcement Learning for Task Specifications with Action-Constraints2 Jan 2022 0 repositories listed
-
Robust Algorithmic Collusion2 Jan 2022 0 repositories listed
-
A Surrogate-Assisted Controller for Expensive Evolutionary Reinforcement Learning1 Jan 2022 0 repositories listed
-
Joint Learning-Based Stabilization of Multiple Unknown Linear Systems1 Jan 2022 0 repositories listed
-
Operator Deep Q-Learning: Zero-Shot Reward Transferring in Reinforcement Learning1 Jan 2022 0 repositories listed
-
Symmetry-Aware Neural Architecture for Embodied Visual Exploration1 Jan 2022 0 repositories listed
-
Temporal Complementarity-Guided Reinforcement Learning for Image-to-Video Person Re-Identification1 Jan 2022 0 repositories listed
-
Toward Pareto Efficient Fairness-Utility Trade-off inRecommendation through Reinforcement Learning1 Jan 2022 0 repositories listed
-
Transfer RL across Observation Feature Spaces via Model-Based Regularization1 Jan 2022 0 repositories listed
-
Importance of Empirical Sample Complexity Analysis for Offline Reinforcement Learning31 Dec 2021 0 repositories listed
-
Robust Entropy-regularized Markov Decision Processes31 Dec 2021 0 repositories listed
-
Single-Shot Pruning for Offline Reinforcement Learning31 Dec 2021 0 repositories listed
-
Stochastic convex optimization for provably efficient apprenticeship learning31 Dec 2021 0 repositories listed
-
Using Graph-Aware Reinforcement Learning to Identify Winning Strategies in Diplomacy Games (Student Abstract)31 Dec 2021 0 repositories listed
-
Constructing a Good Behavior Basis for Transfer using Generalized Policy Updates30 Dec 2021 0 repositories listed
-
Multi-Agent Reinforcement Learning via Adaptive Kalman Temporal Difference and Successor Representation30 Dec 2021 0 repositories listed
-
Reversible Upper Confidence Bound Algorithm to Generate Diverse Optimized Candidates30 Dec 2021 0 repositories listed
-
Stability-Preserving Automatic Tuning of PID Control with Reinforcement Learning30 Dec 2021 0 repositories listed
-
Control Theoretic Analysis of Temporal Difference Learning29 Dec 2021 0 repositories listed
-
Modified DDPG car-following model with a real-world human driving experience with CARLA simulator29 Dec 2021 0 repositories listed
-
Efficient Performance Bounds for Primal-Dual Reinforcement Learning from Demonstrations28 Dec 2021 0 repositories listed
-
Embodied Learning for Lifelong Visual Perception28 Dec 2021 0 repositories listed
-
Robustness and risk management via distributional dynamic programming28 Dec 2021 0 repositories listed
-
Multiagent Model-based Credit Assignment for Continuous Control27 Dec 2021 0 repositories listed
-
A Graph Attention Learning Approach to Antenna Tilt Optimization27 Dec 2021 0 repositories listed
-
Can Reinforcement Learning Find Stackelberg-Nash Equilibria in General-Sum Markov Games with Myopic Followers?27 Dec 2021 0 repositories listed
-
RELDEC: Reinforcement Learning-Based Decoding of Moderate Length LDPC Codes27 Dec 2021 0 repositories listed
-
Safe Reinforcement Learning with Chance-constrained Model Predictive Control27 Dec 2021 0 repositories listed
-
The Statistical Complexity of Interactive Decision Making27 Dec 2021 0 repositories listed
-
Abstractions of General Reinforcement Learning26 Dec 2021 0 repositories listed
-
Neuro-Symbolic Hierarchical Rule Induction26 Dec 2021 0 repositories listed
-
Reducing Planning Complexity of General Reinforcement Learning with Non-Markovian Abstractions26 Dec 2021 0 repositories listed
-
A Survey on Interpretable Reinforcement Learning24 Dec 2021 0 repositories listed
-
Dynamic Channel Access via Meta-Reinforcement Learning24 Dec 2021 0 repositories listed
-
Rediscovering Affordance: A Reinforcement Learning Perspective24 Dec 2021 0 repositories listed
-
Improving the Efficiency of Off-Policy Reinforcement Learning by Accounting for Past Decisions23 Dec 2021 0 repositories listed
-
Local Advantage Networks for Cooperative Multi-Agent Reinforcement Learning23 Dec 2021 0 repositories listed
-
Missing Velocity in Dynamic Obstacle Avoidance based on Deep Reinforcement Learning23 Dec 2021 0 repositories listed
-
Deep Reinforcement Learning for Optimal Power Flow with Renewables Using Graph Information22 Dec 2021 0 repositories listed
-
Graph augmented Deep Reinforcement Learning in the GameRLand3D environment22 Dec 2021 0 repositories listed
-
A Scalable Deep Reinforcement Learning Model for Online Scheduling Coflows of Multi-Stage Jobs for High Performance Computing21 Dec 2021 0 repositories listed
-
Aerial Base Station Positioning and Power Control for Securing Communications: A Deep Q-Network Approach21 Dec 2021 0 repositories listed
-
District Cooling System Control for Providing Operating Reserve based on Safe Deep Reinforcement Learning21 Dec 2021 0 repositories listed
-
Do Androids Dream of Electric Fences? Safety-Aware Reinforcement Learning with Latent Shielding21 Dec 2021 0 repositories listed
-
Nearly Optimal Policy Optimization with Stable at Any Time Guarantee21 Dec 2021 0 repositories listed
-
Reinforcement Learning based Sequential Batch-sampling for Bayesian Optimal Experimental Design21 Dec 2021 0 repositories listed
-
A deep reinforcement learning model for predictive maintenance planning of road assets: Integrating LCA and LCCA20 Dec 2021 0 repositories listed
-
AGPNet -- Autonomous Grading Policy Network20 Dec 2021 0 repositories listed
-
Interpretable Preference-based Reinforcement Learning with Tree-Structured Reward Functions20 Dec 2021 0 repositories listed