Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 116
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 116 of 152: papers 11,501 to 11,600 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Sentiment Analysis for Reinforcement Learning5 Oct 2020 0 repositories listed
-
The act of remembering: a study in partially observable reinforcement learning5 Oct 2020 0 repositories listed
-
A Sharp Analysis of Model-based Reinforcement Learning with Self-Play4 Oct 2020 0 repositories listed
-
Test-Cost Sensitive Methods for Identifying Nearby Points4 Oct 2020 0 repositories listed
-
Attractor Selection in Nonlinear Energy Harvesting Using Deep Reinforcement Learning3 Oct 2020 0 repositories listed
-
Beyond Tabula-Rasa: a Modular Reinforcement Learning Approach for Physically Embedded 3D Sokoban3 Oct 2020 0 repositories listed
-
Disentangling causal effects for hierarchical reinforcement learning3 Oct 2020 0 repositories listed
-
Mean-Variance Efficient Reinforcement Learning with Applications to Dynamic Financial Investment3 Oct 2020 0 repositories listed
-
Interactive Reinforcement Learning for Feature Selection with Decision Tree in the Loop2 Oct 2020 0 repositories listed
-
MADRaS : Multi Agent Driving Simulator2 Oct 2020 0 repositories listed
-
Reinforcement Learning of Sequential Price Mechanisms2 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning with Mixed Convolutional Network1 Oct 2020 0 repositories listed
-
Nearly Minimax Optimal Reinforcement Learning for Discounted MDPs1 Oct 2020 0 repositories listed
-
Multi-Reward based Reinforcement Learning for Neural Machine Translation1 Oct 2020 0 repositories listed
-
Recognition Method of Important Words in Korean Text based on Reinforcement Learning1 Oct 2020 0 repositories listed
-
Bayesian Meta-reinforcement Learning for Traffic Signal Control1 Oct 2020 0 repositories listed
-
AAMDRL: Augmented Asset Management with Deep Reinforcement Learning30 Sep 2020 0 repositories listed
-
Accelerating Optimization and Reinforcement Learning with Quasi-Stochastic Approximation30 Sep 2020 0 repositories listed
-
Bridging the gap between Markowitz planning and deep reinforcement learning30 Sep 2020 0 repositories listed
-
Entropy Regularization for Mean Field Games with Learning30 Sep 2020 0 repositories listed
-
Finding It at Another Side: A Viewpoint-Adapted Matching Encoder for Change Captioning30 Sep 2020 0 repositories listed
-
Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network30 Sep 2020 0 repositories listed
-
Strategy and Benchmark for Converting Deep Q-Networks to Event-Driven Spiking Neural Networks30 Sep 2020 0 repositories listed
-
Teacher-Critical Training Strategies for Image Captioning30 Sep 2020 0 repositories listed
-
Toolpath design for additive manufacturing using deep reinforcement learning30 Sep 2020 0 repositories listed
-
Cross Learning in Deep Q-Networks29 Sep 2020 0 repositories listed
-
Trust-Region Method with Deep Reinforcement Learning in Analog Design Space Exploration29 Sep 2020 0 repositories listed
-
Multi-objective Reinforcement Learning based approach for User-Centric Power Optimization in Smart Home Environments29 Sep 2020 0 repositories listed
-
Reannealing of Decaying Exploration Based On Heuristic Measure in Deep Q-Network29 Sep 2020 0 repositories listed
-
Agent Environment Cycle Games28 Sep 2020 0 repositories listed
-
Deep Reinforcement Learning for DER Cyber-Attack Mitigation28 Sep 2020 0 repositories listed
-
Efficient Exploration for Model-based Reinforcement Learning with Continuous States and Actions28 Sep 2020 0 repositories listed
-
Is Reinforcement Learning More Difficult Than Bandits? A Near-optimal Algorithm Escaping the Curse of Horizon28 Sep 2020 0 repositories listed
-
Jointly-Trained State-Action Embedding for Efficient Reinforcement Learning28 Sep 2020 0 repositories listed
-
MDP Playground: Controlling Orthogonal Dimensions of Hardness in Toy Environments28 Sep 2020 0 repositories listed
-
Near-Optimal Regret Bounds for Model-Free RL in Non-Stationary Episodic MDPs28 Sep 2020 0 repositories listed
-
Neuron Activation Analysis for Multi-Joint Robot Reinforcement Learning28 Sep 2020 0 repositories listed
-
Policy Gradient with Expected Quadratic Utility Maximization: A New Mean-Variance Approach in Reinforcement Learning28 Sep 2020 0 repositories listed
-
REPAINT: Knowledge Transfer in Deep Actor-Critic Reinforcement Learning28 Sep 2020 0 repositories listed
-
The Emergence of Individuality in Multi-Agent Reinforcement Learning28 Sep 2020 0 repositories listed
-
Towards Heterogeneous Multi-Agent Reinforcement Learning with Graph Neural Networks28 Sep 2020 0 repositories listed
-
Transfer among Agents: An Efficient Multiagent Transfer Learning Framework28 Sep 2020 0 repositories listed
-
What About Taking Policy as Input of Value Function: Policy-extended Value Function Approximator28 Sep 2020 0 repositories listed
-
Machine Learning in Event-Triggered Control: Recent Advances and Open Issues27 Sep 2020 0 repositories listed
-
Scalable Deep Reinforcement Learning for Ride-Hailing27 Sep 2020 0 repositories listed
-
Scheduling and Power Control for Wireless Multicast Systems via Deep Reinforcement Learning27 Sep 2020 0 repositories listed
-
Virtual Experience to Real World Application: Sidewalk Obstacle Avoidance Using Reinforcement Learning for Visually Impaired27 Sep 2020 0 repositories listed
-
Complementary Meta-Reinforcement Learning for Fault-Adaptive Control26 Sep 2020 0 repositories listed
-
Graph neural induction of value iteration26 Sep 2020 0 repositories listed
-
Inverse Rational Control with Partially Observable Continuous Nonlinear Dynamics26 Sep 2020 0 repositories listed
-
Lineage Evolution Reinforcement Learning26 Sep 2020 0 repositories listed
-
Reinforcement Learning-based N-ary Cross-Sentence Relation Extraction26 Sep 2020 0 repositories listed
-
Motion Planning by Reinforcement Learning for an Unmanned Aerial Vehicle in Virtual Open Space with Static Obstacles24 Sep 2020 0 repositories listed
-
Sim-to-Real Transfer in Deep Reinforcement Learning for Robotics: a Survey24 Sep 2020 0 repositories listed
-
A Multi-Agent Deep Reinforcement Learning Approach for a Distributed Energy Marketplace in Smart Grids23 Sep 2020 0 repositories listed
-
Demand Responsive Dynamic Pricing Framework for Prosumer Dominated Microgrids using Multiagent Reinforcement Learning23 Sep 2020 0 repositories listed
-
Probabilistic Machine Learning for Healthcare23 Sep 2020 0 repositories listed
-
ReLeaSER: A Reinforcement Learning Strategy for Optimizing Utilization Of Ephemeral Cloud Resources23 Sep 2020 0 repositories listed
-
Robust Reinforcement Learning-based Autonomous Driving Agent for Simulation and Real World23 Sep 2020 0 repositories listed
-
What is the Reward for Handwriting? -- Handwriting Generation by Imitation Learning23 Sep 2020 0 repositories listed
-
Deep Reinforcement Learning for On-line Dialogue State Tracking22 Sep 2020 0 repositories listed
-
Distributed Structured Actor-Critic Reinforcement Learning for Universal Dialogue Management22 Sep 2020 0 repositories listed
-
Is Q-Learning Provably Efficient? An Extended Analysis22 Sep 2020 0 repositories listed
-
SUMBT+LaRL: Effective Multi-domain End-to-end Neural Task-oriented Dialog System22 Sep 2020 0 repositories listed
-
Contextual Bandits for adapting to changing User preferences over time21 Sep 2020 0 repositories listed
-
DISPATCH: Design Space Exploration of Cyber-Physical Systems21 Sep 2020 0 repositories listed
-
Dynamic Horizon Value Estimation for Model-based Reinforcement Learning21 Sep 2020 0 repositories listed
-
Human Engagement Providing Evaluative and Informative Advice for Interactive Reinforcement Learning21 Sep 2020 0 repositories listed
-
Learn to Exceed: Stereo Inverse Reinforcement Learning with Concurrent Policy Optimization21 Sep 2020 0 repositories listed
-
Learning a Contact-Adaptive Controller for Robust, Efficient Legged Locomotion21 Sep 2020 0 repositories listed
-
Mobile Cellular-Connected UAVs: Reinforcement Learning for Sky Limits21 Sep 2020 0 repositories listed
-
Reinforcement Learning Approaches in Social Robotics21 Sep 2020 0 repositories listed
-
Lyapunov-Based Reinforcement Learning for Decentralized Multi-Agent Control20 Sep 2020 0 repositories listed
-
Multiplayer Support for the Arcade Learning Environment20 Sep 2020 0 repositories listed
-
Regret Bounds and Reinforcement Learning Exploration of EXP-based Algorithms20 Sep 2020 0 repositories listed
-
Construction of Polar Codes with Reinforcement Learning19 Sep 2020 0 repositories listed
-
A Contraction Approach to Model-based Reinforcement Learning18 Sep 2020 0 repositories listed
-
Deep Reinforcement Learning for Closed-Loop Blood Glucose Control18 Sep 2020 0 repositories listed
-
Private Reinforcement Learning with PAC and Regret Guarantees18 Sep 2020 0 repositories listed
-
Reinforcement Learning for Weakly Supervised Temporal Grounding of Natural Language in Untrimmed Videos18 Sep 2020 0 repositories listed
-
GeneraLight: Improving Environment Generalization of Traffic Signal Control via Meta Reinforcement Learning17 Sep 2020 0 repositories listed
-
Knowledge-Assisted Deep Reinforcement Learning in 5G Scheduler Design: From Theoretical Framework to Implementation17 Sep 2020 0 repositories listed
-
Reward Maximisation through Discrete Active Inference17 Sep 2020 0 repositories listed
-
Reconstructing Actions To Explain Deep Reinforcement Learning17 Sep 2020 0 repositories listed
-
DRL-FAS: A Novel Framework Based on Deep Reinforcement Learning for Face Anti-Spoofing16 Sep 2020 0 repositories listed
-
Theory of Mind with Guilt Aversion Facilitates Cooperative Reinforcement Learning16 Sep 2020 0 repositories listed
-
Time your hedge with Deep Reinforcement Learning16 Sep 2020 0 repositories listed
-
Transfer Learning in Deep Reinforcement Learning: A Survey16 Sep 2020 0 repositories listed
-
Autonomous Learning of Features for Control: Experiments with Embodied and Situated Agents15 Sep 2020 0 repositories listed
-
Decoding Polar Codes with Reinforcement Learning15 Sep 2020 0 repositories listed
-
Reinforcement Learning for Strategic Recommendations15 Sep 2020 0 repositories listed
-
Soft policy optimization using dual-track advantage estimator15 Sep 2020 0 repositories listed
-
Efficient Transformers: A Survey14 Sep 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning in Cournot Games14 Sep 2020 0 repositories listed
-
Predictive Synthesis of Quantum Materials by Probabilistic Reinforcement Learning14 Sep 2020 0 repositories listed
-
Reinforcement Learning for Dynamic Resource Optimization in 5G Radio Access Network Slicing14 Sep 2020 0 repositories listed
-
Variance-Reduced Off-Policy Memory-Efficient Policy Search14 Sep 2020 0 repositories listed
-
Efficient Competitive Self-Play Policy Optimization13 Sep 2020 0 repositories listed
-
Extended Radial Basis Function Controller for Reinforcement Learning12 Sep 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.