Browse State-of-the-Art › Reinforcement Learning › Papers, page 90
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 90 of 132: papers 8,901 to 9,000 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Reinforcement Learning with Efficient Active Feature Acquisition2 Nov 2020 0 repositories listed
-
Sample-efficient reinforcement learning using deep Gaussian processes2 Nov 2020 0 repositories listed
-
Reinforcement Learning with Imbalanced Dataset for Data-to-Text Medical Report Generation1 Nov 2020 0 repositories listed
-
Abstract Value Iteration for Hierarchical Reinforcement Learning29 Oct 2020 0 repositories listed
-
Reinforcement Learning of Causal Variables Using Mediation Analysis29 Oct 2020 0 repositories listed
-
How do Offline Measures for Exploration in Reinforcement Learning behave?29 Oct 2020 0 repositories listed
-
Machine versus Human Attention in Deep Reinforcement Learning Tasks29 Oct 2020 0 repositories listed
-
DeepFoldit -- A Deep Reinforcement Learning Neural Network Folding Proteins28 Oct 2020 0 repositories listed
-
Designing Interpretable Approximations to Deep Reinforcement Learning28 Oct 2020 0 repositories listed
-
Batch Reinforcement Learning with a Nonparametric Off-Policy Policy Gradient27 Oct 2020 0 repositories listed
-
Behavior Priors for Efficient Reinforcement Learning27 Oct 2020 0 repositories listed
-
Can Reinforcement Learning for Continuous Control Generalize Across Physics Engines?27 Oct 2020 0 repositories listed
-
Behavioral decision-making for urban autonomous driving in the presence of pedestrians using Deep Recurrent Q-Network26 Oct 2020 0 repositories listed
-
Lyapunov-Based Reinforcement Learning State Estimator26 Oct 2020 0 repositories listed
-
OPAL: Offline Primitive Discovery for Accelerating Offline Reinforcement Learning26 Oct 2020 0 repositories listed
-
Pairwise heuristic sequence alignment algorithm based on deep reinforcement learning26 Oct 2020 0 repositories listed
-
VisualHints: A Visual-Lingual Environment for Multimodal Reinforcement Learning26 Oct 2020 0 repositories listed
-
Improving the Exploration of Deep Reinforcement Learning in Continuous Domains using Planning for Policy Search24 Oct 2020 0 repositories listed
-
Planning with Exploration: Addressing Dynamics Bottleneck in Model-based Reinforcement Learning24 Oct 2020 0 repositories listed
-
Improved Worst-Case Regret Bounds for Randomized Least-Squares Value Iteration23 Oct 2020 0 repositories listed
-
Option Hedging with Risk Averse Reinforcement Learning23 Oct 2020 0 repositories listed
-
Stochastic Inverse Reinforcement Learning23 Oct 2020 0 repositories listed
-
Error Bounds of Imitating Policies and Environments22 Oct 2020 0 repositories listed
-
Optimising Stochastic Routing for Taxi Fleets with Model Enhanced Reinforcement Learning22 Oct 2020 0 repositories listed
-
Sample Efficient Reinforcement Learning with REINFORCE22 Oct 2020 0 repositories listed
-
What are the Statistical Limits of Offline RL with Linear Function Approximation?22 Oct 2020 0 repositories listed
-
Safety Verification of Model Based Reinforcement Learning Controllers21 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning in Lane Merge Coordination for Connected Vehicles20 Oct 2020 0 repositories listed
-
Language Inference with Multi-head Automata through Reinforcement Learning20 Oct 2020 0 repositories listed
-
Negotiating Team Formation Using Deep Reinforcement Learning20 Oct 2020 0 repositories listed
-
Quality of service based radar resource management using deep reinforcement learning20 Oct 2020 0 repositories listed
-
Robust Constrained Reinforcement Learning for Continuous Control with Model Misspecification20 Oct 2020 0 repositories listed
-
Runtime Safety Assurance Using Reinforcement Learning20 Oct 2020 0 repositories listed
-
A Reinforcement Learning Approach to Health Aware Control Strategy19 Oct 2020 0 repositories listed
-
Chance-Constrained Control with Lexicographic Deep Reinforcement Learning19 Oct 2020 0 repositories listed
-
Imitation with Neural Density Models19 Oct 2020 0 repositories listed
-
Average-reward model-free reinforcement learning: a systematic review and literature mapping18 Oct 2020 0 repositories listed
-
Model-Based Inverse Reinforcement Learning from Visual Demonstrations18 Oct 2020 0 repositories listed
-
Assessment of Reward Functions in Reinforcement Learning for Multi-Modal Urban Traffic Control under Real-World limitations17 Oct 2020 0 repositories listed
-
Learning Elimination Ordering for Tree Decomposition Problem17 Oct 2020 0 repositories listed
-
Learning Lower Bounds for Graph Exploration With Reinforcement Learning17 Oct 2020 0 repositories listed
-
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning17 Oct 2020 0 repositories listed
-
Autonomous Control of a Particle Accelerator using Deep Reinforcement Learning16 Oct 2020 0 repositories listed
-
DOOM: A Novel Adversarial-DRL-Based Op-Code Level Metamorphic Malware Obfuscator for the Enhancement of IDS16 Oct 2020 0 repositories listed
-
Efficient Robotic Object Search via HIEM: Hierarchical Policy Learning with Intrinsic-Extrinsic Modeling16 Oct 2020 0 repositories listed
-
Reinforcement Learning for Efficient and Tuning-Free Link Adaptation16 Oct 2020 0 repositories listed
-
Uncertainty-aware Contact-safe Model-based Reinforcement Learning16 Oct 2020 0 repositories listed
-
A Nesterov's Accelerated quasi-Newton method for Global Routing using Deep Reinforcement Learning15 Oct 2020 0 repositories listed
-
An Empowerment-based Solution to Robotic Manipulation Tasks with Sparse Rewards15 Oct 2020 0 repositories listed
-
Blending Search and Discovery: Tag-Based Query Refinement with Contextual Reinforcement Learning15 Oct 2020 0 repositories listed
-
Cooperative-Competitive Reinforcement Learning with History-Dependent Rewards15 Oct 2020 0 repositories listed
-
Explanation Augmented Feedback in Human-in-the-Loop Reinforcement Learning15 Oct 2020 0 repositories listed
-
Local Differential Privacy for Regret Minimization in Reinforcement Learning15 Oct 2020 0 repositories listed
-
Optimal Dispatch in Emergency Service System via Reinforcement Learning15 Oct 2020 0 repositories listed
-
Average Cost Optimal Control of Stochastic Systems Using Reinforcement Learning13 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning and Transportation Research: A Comprehensive Review13 Oct 2020 0 repositories listed
-
Model-Based Reinforcement Learning for Type 1Diabetes Blood Glucose Control13 Oct 2020 0 repositories listed
-
12 Oct 2020 0 repositories listed
-
Nearly Minimax Optimal Reward-free Reinforcement Learning12 Oct 2020 0 repositories listed
-
Remote Electrical Tilt Optimization via Safe Reinforcement Learning12 Oct 2020 0 repositories listed
-
Controlling Graph Dynamics with Reinforcement Learning and Graph Neural Networks11 Oct 2020 0 repositories listed
-
Safe Reinforcement Learning with Natural Language Constraints11 Oct 2020 0 repositories listed
-
Reinforcement Learning on Computational Resource Allocation of Cloud-based Wireless Networks10 Oct 2020 0 repositories listed
-
Characterizing Policy Divergence for Personalized Meta-Reinforcement Learning9 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Asset Allocation in US Equities9 Oct 2020 0 repositories listed
-
Deep RL With Information Constrained Policies: Generalization in Continuous Control9 Oct 2020 0 repositories listed
-
Jointly-Learned State-Action Embedding for Efficient Reinforcement Learning9 Oct 2020 0 repositories listed
-
Parameterized Reinforcement Learning for Optical System Optimization9 Oct 2020 0 repositories listed
-
Learning Intrinsic Symbolic Rewards in Reinforcement Learning8 Oct 2020 0 repositories listed
-
Nonstationary Reinforcement Learning with Linear Function Approximation8 Oct 2020 0 repositories listed
-
Provable Fictitious Play for General Mean-Field Games8 Oct 2020 0 repositories listed
-
Episodic Reinforcement Learning in Finite MDPs: Minimax Lower Bounds Revisited7 Oct 2020 0 repositories listed
-
Instance-Dependent Complexity of Contextual Bandits and Reinforcement Learning: A Disagreement-Based Perspective7 Oct 2020 0 repositories listed
-
Online Safety Assurance for Deep Reinforcement Learning7 Oct 2020 0 repositories listed
-
Regularized Inverse Reinforcement Learning7 Oct 2020 0 repositories listed
-
Heterogeneous Multi-Agent Reinforcement Learning for Unknown Environment Mapping6 Oct 2020 0 repositories listed
-
Safety Aware Reinforcement Learning (SARL)6 Oct 2020 0 repositories listed
-
UneVEn: Universal Value Exploration for Multi-Agent Reinforcement Learning6 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Collaborative Edge Computing in Vehicular Networks5 Oct 2020 0 repositories listed
-
Goal-directed Generation of Discrete Structures with Conditional Generative Models5 Oct 2020 0 repositories listed
-
Sentiment Analysis for Reinforcement Learning5 Oct 2020 0 repositories listed
-
A Sharp Analysis of Model-based Reinforcement Learning with Self-Play4 Oct 2020 0 repositories listed
-
Attractor Selection in Nonlinear Energy Harvesting Using Deep Reinforcement Learning3 Oct 2020 0 repositories listed
-
Disentangling causal effects for hierarchical reinforcement learning3 Oct 2020 0 repositories listed
-
Mean-Variance Efficient Reinforcement Learning with Applications to Dynamic Financial Investment3 Oct 2020 0 repositories listed
-
Reinforcement Learning of Sequential Price Mechanisms2 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning with Mixed Convolutional Network1 Oct 2020 0 repositories listed
-
Nearly Minimax Optimal Reinforcement Learning for Discounted MDPs1 Oct 2020 0 repositories listed
-
Multi-Reward based Reinforcement Learning for Neural Machine Translation1 Oct 2020 0 repositories listed
-
Recognition Method of Important Words in Korean Text based on Reinforcement Learning1 Oct 2020 0 repositories listed
-
Bayesian Meta-reinforcement Learning for Traffic Signal Control1 Oct 2020 0 repositories listed
-
AAMDRL: Augmented Asset Management with Deep Reinforcement Learning30 Sep 2020 0 repositories listed
-
Accelerating Optimization and Reinforcement Learning with Quasi-Stochastic Approximation30 Sep 2020 0 repositories listed
-
Bridging the gap between Markowitz planning and deep reinforcement learning30 Sep 2020 0 repositories listed
-
Graph-based Heuristic Search for Module Selection Procedure in Neural Module Network30 Sep 2020 0 repositories listed
-
Strategy and Benchmark for Converting Deep Q-Networks to Event-Driven Spiking Neural Networks30 Sep 2020 0 repositories listed
-
Toolpath design for additive manufacturing using deep reinforcement learning30 Sep 2020 0 repositories listed
-
Deep Reinforcement Learning for DER Cyber-Attack Mitigation28 Sep 2020 0 repositories listed
-
Is Reinforcement Learning More Difficult Than Bandits? A Near-optimal Algorithm Escaping the Curse of Horizon28 Sep 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.