Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 115
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 115 of 152: papers 11,401 to 11,500 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Option Hedging with Risk Averse Reinforcement Learning23 Oct 2020 0 repositories listed
-
Stabilizing Transformer-Based Action Sequence Generation For Q-Learning23 Oct 2020 0 repositories listed
-
Stochastic Inverse Reinforcement Learning23 Oct 2020 0 repositories listed
-
Adversarial Attacks on Deep Algorithmic Trading Policies22 Oct 2020 0 repositories listed
-
CoinDICE: Off-Policy Confidence Interval Estimation22 Oct 2020 0 repositories listed
-
Error Bounds of Imitating Policies and Environments22 Oct 2020 0 repositories listed
-
Incorporating Stylistic Lexical Preferences in Generative Language Models22 Oct 2020 0 repositories listed
-
Motion Planner Augmented Reinforcement Learning for Robot Manipulation in Obstructed Environments22 Oct 2020 0 repositories listed
-
Optimising Stochastic Routing for Taxi Fleets with Model Enhanced Reinforcement Learning22 Oct 2020 0 repositories listed
-
Optimizing Coverage and Capacity in Cellular Networks using Machine Learning22 Oct 2020 0 repositories listed
-
Sample Efficient Reinforcement Learning with REINFORCE22 Oct 2020 0 repositories listed
-
What are the Statistical Limits of Offline RL with Linear Function Approximation?22 Oct 2020 0 repositories listed
-
Logistic Q-Learning21 Oct 2020 0 repositories listed
-
On Information Asymmetry in Competitive Multi-Agent Reinforcement Learning: Convergence and Optimality21 Oct 2020 0 repositories listed
-
Reinforcement learning using Deep Q Networks and Q learning accurately localizes brain tumors on MRI with very small training sets21 Oct 2020 0 repositories listed
-
Safety Verification of Model Based Reinforcement Learning Controllers21 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning in Lane Merge Coordination for Connected Vehicles20 Oct 2020 0 repositories listed
-
Integrating LEO Satellites and Multi-UAV Reinforcement Learning for Hybrid FSO/RF Non-Terrestrial Networks20 Oct 2020 0 repositories listed
-
Language Inference with Multi-head Automata through Reinforcement Learning20 Oct 2020 0 repositories listed
-
Multi-Radar Tracking Optimization for Collaborative Combat20 Oct 2020 0 repositories listed
-
Negotiating Team Formation Using Deep Reinforcement Learning20 Oct 2020 0 repositories listed
-
Quality of service based radar resource management using deep reinforcement learning20 Oct 2020 0 repositories listed
-
Robust Constrained Reinforcement Learning for Continuous Control with Model Misspecification20 Oct 2020 0 repositories listed
-
Runtime Safety Assurance Using Reinforcement Learning20 Oct 2020 0 repositories listed
-
A case for new neural networks smoothness constraints19 Oct 2020 0 repositories listed
-
A Reinforcement Learning Approach to Health Aware Control Strategy19 Oct 2020 0 repositories listed
-
Chance-Constrained Control with Lexicographic Deep Reinforcement Learning19 Oct 2020 0 repositories listed
-
Evaluating the Safety of Deep Reinforcement Learning Models using Semi-Formal Verification19 Oct 2020 0 repositories listed
-
Imitation with Neural Density Models19 Oct 2020 0 repositories listed
-
Average-reward model-free reinforcement learning: a systematic review and literature mapping18 Oct 2020 0 repositories listed
-
Model-Based Inverse Reinforcement Learning from Visual Demonstrations18 Oct 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning in NOMA-aided UAV Networks for Cellular Offloading18 Oct 2020 0 repositories listed
-
Assessment of Reward Functions in Reinforcement Learning for Multi-Modal Urban Traffic Control under Real-World limitations17 Oct 2020 0 repositories listed
-
Learning Elimination Ordering for Tree Decomposition Problem17 Oct 2020 0 repositories listed
-
Learning Lower Bounds for Graph Exploration With Reinforcement Learning17 Oct 2020 0 repositories listed
-
Neural Algorithms for Graph Navigation17 Oct 2020 0 repositories listed
-
Scalable Evolution Strategies Pipeline for Solving the Vehicle Routing Problem17 Oct 2020 0 repositories listed
-
Variational Dynamic for Self-Supervised Exploration in Deep Reinforcement Learning17 Oct 2020 0 repositories listed
-
Autonomous Control of a Particle Accelerator using Deep Reinforcement Learning16 Oct 2020 0 repositories listed
-
Collaborative Training of GANs in Continuous and Discrete Spaces for Text Generation16 Oct 2020 0 repositories listed
-
DOOM: A Novel Adversarial-DRL-Based Op-Code Level Metamorphic Malware Obfuscator for the Enhancement of IDS16 Oct 2020 0 repositories listed
-
Efficient Robotic Object Search via HIEM: Hierarchical Policy Learning with Intrinsic-Extrinsic Modeling16 Oct 2020 0 repositories listed
-
Few-shot model-based adaptation in noisy conditions16 Oct 2020 0 repositories listed
-
Decomposability and Parallel Computation of Multi-Agent LQR16 Oct 2020 0 repositories listed
-
Interpretable Disease Prediction based on Reinforcement Path Reasoning over Knowledge Graphs16 Oct 2020 0 repositories listed
-
Reinforcement Learning for Efficient and Tuning-Free Link Adaptation16 Oct 2020 0 repositories listed
-
Uncertainty-aware Contact-safe Model-based Reinforcement Learning16 Oct 2020 0 repositories listed
-
A Nesterov's Accelerated quasi-Newton method for Global Routing using Deep Reinforcement Learning15 Oct 2020 0 repositories listed
-
An Empowerment-based Solution to Robotic Manipulation Tasks with Sparse Rewards15 Oct 2020 0 repositories listed
-
Applicability and Challenges of Deep Reinforcement Learning for Satellite Frequency Plan Design15 Oct 2020 0 repositories listed
-
Blending Search and Discovery: Tag-Based Query Refinement with Contextual Reinforcement Learning15 Oct 2020 0 repositories listed
-
Cooperative-Competitive Reinforcement Learning with History-Dependent Rewards15 Oct 2020 0 repositories listed
-
Deep Learning of Koopman Representation for Control15 Oct 2020 0 repositories listed
-
Explanation Augmented Feedback in Human-in-the-Loop Reinforcement Learning15 Oct 2020 0 repositories listed
-
Local Differential Privacy for Regret Minimization in Reinforcement Learning15 Oct 2020 0 repositories listed
-
Optimal Dispatch in Emergency Service System via Reinforcement Learning15 Oct 2020 0 repositories listed
-
Average Cost Optimal Control of Stochastic Systems Using Reinforcement Learning13 Oct 2020 0 repositories listed
-
Balancing Constraints and Rewards with Meta-Gradient D4PG13 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning and Transportation Research: A Comprehensive Review13 Oct 2020 0 repositories listed
-
Grid-Interactive Multi-Zone Building Control Using Reinforcement Learning with Global-Local Policy Search13 Oct 2020 0 repositories listed
-
Model-Based Reinforcement Learning for Type 1Diabetes Blood Glucose Control13 Oct 2020 0 repositories listed
-
Random Network Distillation as a Diversity Metric for Both Image and Text Generation13 Oct 2020 0 repositories listed
-
AttendLight: Universal Attention-Based Reinforcement Learning Model for Traffic Signal Control12 Oct 2020 0 repositories listed
-
12 Oct 2020 0 repositories listed
-
Is Plug-in Solver Sample-Efficient for Feature-based Reinforcement Learning?12 Oct 2020 0 repositories listed
-
Local Search for Policy Iteration in Continuous Control12 Oct 2020 0 repositories listed
-
Nearly Minimax Optimal Reward-free Reinforcement Learning12 Oct 2020 0 repositories listed
-
Remote Electrical Tilt Optimization via Safe Reinforcement Learning12 Oct 2020 0 repositories listed
-
The Greatest Teacher, Failure is: Using Reinforcement Learning for SFC Placement Based on Availability and Energy Consumption12 Oct 2020 0 repositories listed
-
Deep-Reinforcement-Learning-Based Scheduling with Contiguous Resource Allocation for Next-Generation Cellular Systems11 Oct 2020 0 repositories listed
-
Controlling Graph Dynamics with Reinforcement Learning and Graph Neural Networks11 Oct 2020 0 repositories listed
-
Safe Reinforcement Learning with Natural Language Constraints11 Oct 2020 0 repositories listed
-
MS-Ranker: Accumulating Evidence from Potentially Correct Candidates for Answer Selection10 Oct 2020 0 repositories listed
-
Reinforcement Learning on Computational Resource Allocation of Cloud-based Wireless Networks10 Oct 2020 0 repositories listed
-
Trust the Model When It Is Confident: Masked Model-based Actor-Critic10 Oct 2020 0 repositories listed
-
Characterizing Policy Divergence for Personalized Meta-Reinforcement Learning9 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Asset Allocation in US Equities9 Oct 2020 0 repositories listed
-
Deep RL With Information Constrained Policies: Generalization in Continuous Control9 Oct 2020 0 repositories listed
-
Dynamic Context Selection for Document-level Neural Machine Translation via Reinforcement Learning9 Oct 2020 0 repositories listed
-
Jointly-Learned State-Action Embedding for Efficient Reinforcement Learning9 Oct 2020 0 repositories listed
-
Learning to Locomote: Understanding How Environment Design Matters for Deep Reinforcement Learning9 Oct 2020 0 repositories listed
-
Parameterized Reinforcement Learning for Optical System Optimization9 Oct 2020 0 repositories listed
-
Learning Intrinsic Symbolic Rewards in Reinforcement Learning8 Oct 2020 0 repositories listed
-
Nonstationary Reinforcement Learning with Linear Function Approximation8 Oct 2020 0 repositories listed
-
Provable Fictitious Play for General Mean-Field Games8 Oct 2020 0 repositories listed
-
Actor-Critic Algorithm for High-dimensional Partial Differential Equations7 Oct 2020 0 repositories listed
-
Episodic Reinforcement Learning in Finite MDPs: Minimax Lower Bounds Revisited7 Oct 2020 0 repositories listed
-
Instance-Dependent Complexity of Contextual Bandits and Reinforcement Learning: A Disagreement-Based Perspective7 Oct 2020 0 repositories listed
-
Model-Free Non-Stationary RL: Near-Optimal Regret and Applications in Multi-Agent RL and Inventory Control7 Oct 2020 0 repositories listed
-
Online Safety Assurance for Deep Reinforcement Learning7 Oct 2020 0 repositories listed
-
Regularized Inverse Reinforcement Learning7 Oct 2020 0 repositories listed
-
Reinforcement Learning for Many-Body Ground-State Preparation Inspired by Counterdiabatic Driving7 Oct 2020 0 repositories listed
-
Variational Intrinsic Control Revisited7 Oct 2020 0 repositories listed
-
Heterogeneous Multi-Agent Reinforcement Learning for Unknown Environment Mapping6 Oct 2020 0 repositories listed
-
Safety Aware Reinforcement Learning (SARL)6 Oct 2020 0 repositories listed
-
UneVEn: Universal Value Exploration for Multi-Agent Reinforcement Learning6 Oct 2020 0 repositories listed
-
A Distributed Model-Free Ride-Sharing Approach for Joint Matching, Pricing, and Dispatching using Deep Reinforcement Learning5 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Collaborative Edge Computing in Vehicular Networks5 Oct 2020 0 repositories listed
-
Deep Reinforcement Learning for Electric Vehicle Routing Problem with Time Windows5 Oct 2020 0 repositories listed
-
Goal-directed Generation of Discrete Structures with Conditional Generative Models5 Oct 2020 0 repositories listed