Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 109
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 109 of 152: papers 10,801 to 10,900 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Representation Matters: Offline Pretraining for Sequential Decision Making11 Feb 2021 0 repositories listed
-
Deep Reinforcement Learning with Symmetric Prior for Predictive Power Allocation to Mobile Users10 Feb 2021 0 repositories listed
-
Defense Against Reward Poisoning Attacks in Reinforcement Learning10 Feb 2021 0 repositories listed
-
Derivative-Free Reinforcement Learning: A Review10 Feb 2021 0 repositories listed
-
Learning Equational Theorem Proving10 Feb 2021 0 repositories listed
-
Leveraging Reinforcement Learning for evaluating Robustness of KNN Search Algorithms10 Feb 2021 0 repositories listed
-
Modeling the Interaction between Agents in Cooperative Multi-Agent Reinforcement Learning10 Feb 2021 0 repositories listed
-
Non-stationary Reinforcement Learning without Prior Knowledge: An Optimal Black-box Approach10 Feb 2021 0 repositories listed
-
Patterns, predictions, and actions: A story about machine learning10 Feb 2021 0 repositories listed
-
Personalization for Web-based Services using Offline Reinforcement Learning10 Feb 2021 0 repositories listed
-
Reinforcement Learning for Optimized Beam Training in Multi-Hop Terahertz Communications10 Feb 2021 0 repositories listed
-
Risk-Averse Bayes-Adaptive Reinforcement Learning10 Feb 2021 0 repositories listed
-
Simple Agent, Complex Environment: Efficient Reinforcement Learning with Agent States10 Feb 2021 0 repositories listed
-
Measuring Progress in Deep Reinforcement Learning Sample Efficiency9 Feb 2021 0 repositories listed
-
Adaptive Pairwise Weights for Temporal Credit Assignment9 Feb 2021 0 repositories listed
-
Scheduling the NASA Deep Space Network with Deep Reinforcement Learning9 Feb 2021 0 repositories listed
-
Contrasting Centralized and Decentralized Critics in Multi-Agent Reinforcement Learning8 Feb 2021 0 repositories listed
-
Generate and Revise: Reinforcement Learning in Neural Poetry8 Feb 2021 0 repositories listed
-
Introduction to Machine Learning for the Sciences8 Feb 2021 0 repositories listed
-
Learning Optimal Strategies for Temporal Tasks in Stochastic Games8 Feb 2021 0 repositories listed
-
Provable Model-based Nonlinear Bandit and Reinforcement Learning: Shelve Optimism, Embrace Virtual Curvature8 Feb 2021 0 repositories listed
-
Unlocking Pixels for Reinforcement Learning via Implicit Attention8 Feb 2021 0 repositories listed
-
An Analysis of Frame-skipping in Reinforcement Learning7 Feb 2021 0 repositories listed
-
A bandit approach to curriculum generation for automatic speech recognition6 Feb 2021 0 repositories listed
-
A Hybrid Approach for Reinforcement Learning Using Virtual Policy Gradient for Balancing an Inverted Pendulum6 Feb 2021 0 repositories listed
-
MSPM: A Modularized and Scalable Multi-Agent Reinforcement Learning-based System for Financial Portfolio Management6 Feb 2021 0 repositories listed
-
Improving Model and Search for Computer Go6 Feb 2021 0 repositories listed
-
Multi-Agent Deep Reinforcement Learning for Request Dispatching in Distributed-Controller Software-Defined Networking6 Feb 2021 0 repositories listed
-
Addressing Inherent Uncertainty: Risk-Sensitive Behavior Generation for Automated Driving using Distributional Reinforcement Learning5 Feb 2021 0 repositories listed
-
Deceptive Reinforcement Learning for Privacy-Preserving Planning5 Feb 2021 0 repositories listed
-
Experience-Based Heuristic Search: Robust Motion Planning with Deep Q-Learning5 Feb 2021 0 repositories listed
-
Finite Sample Analysis of Minimax Offline Reinforcement Learning: Completeness, Fast Rates and First-Order Efficiency5 Feb 2021 0 repositories listed
-
Provably Efficient Algorithms for Multi-Objective Competitive RL5 Feb 2021 0 repositories listed
-
A review of motion planning algorithms for intelligent robotics4 Feb 2021 0 repositories listed
-
Deep reinforcement learning-based image classification achieves perfect testing set accuracy for MRI brain tumors with a training set of only 30 images4 Feb 2021 0 repositories listed
-
How to Train Your Robot with Deep Reinforcement Learning; Lessons We've Learned4 Feb 2021 0 repositories listed
-
Hybrid Adversarial Imitation Learning4 Feb 2021 0 repositories listed
-
Persistent Rule-based Interactive Reinforcement Learning4 Feb 2021 0 repositories listed
-
A deep learning model for gas storage optimization3 Feb 2021 0 repositories listed
-
The Pitfall of More Powerful Autoencoders in Lidar-Based Navigation3 Feb 2021 0 repositories listed
-
Multi-UAV Mobile Edge Computing and Path Planning Platform based on Reinforcement Learning3 Feb 2021 0 repositories listed
-
Neural Recursive Belief States in Multi-Agent Reinforcement Learning3 Feb 2021 0 repositories listed
-
A Lyapunov Theory for Finite-Sample Guarantees of Asynchronous Q-Learning and TD-Learning Variants2 Feb 2021 0 repositories listed
-
An Abstraction-based Method to Check Multi-Agent Deep Reinforcement-Learning Behaviors2 Feb 2021 0 repositories listed
-
Approximately Solving Mean Field Games via Entropy-Regularized Deep Reinforcement Learning2 Feb 2021 0 repositories listed
-
Improving Reinforcement Learning with Human Assistance: An Argument for Human Subject Studies with HIPPO Gym2 Feb 2021 0 repositories listed
-
Near-Optimal Offline Reinforcement Learning via Double Variance Reduction2 Feb 2021 0 repositories listed
-
Reinforcement Learning with Probabilistic Boolean Network Models of Smart Grid Devices2 Feb 2021 0 repositories listed
-
A step toward a reinforcement learning de novo genome assembler2 Feb 2021 0 repositories listed
-
Towards Multi-agent Reinforcement Learning for Wireless Network Protocol Synthesis2 Feb 2021 0 repositories listed
-
A Secure Learning Control Strategy via Dynamic Camouflaging for Unknown Dynamical Systems under Attacks1 Feb 2021 0 repositories listed
-
Bellman Eluder Dimension: New Rich Classes of RL Problems, and Sample-Efficient Algorithms1 Feb 2021 0 repositories listed
-
Hybrid Beamforming for mmWave MU-MISO Systems Exploiting Multi-agent Deep Reinforcement Learning1 Feb 2021 0 repositories listed
-
Hybrid Information-driven Multi-agent Reinforcement Learning1 Feb 2021 0 repositories listed
-
Interpretable Reinforcement Learning Inspired by Piaget's Theory of Cognitive Development1 Feb 2021 0 repositories listed
-
Risk Aware and Multi-Objective Decision Making with Distributional Monte Carlo Tree Search1 Feb 2021 0 repositories listed
-
Throughput Optimization for Grant-Free Multiple Access With Multiagent Deep Reinforcement Learning1 Feb 2021 0 repositories listed
-
Fast Rates for the Regret of Offline Reinforcement Learning31 Jan 2021 0 repositories listed
-
Improving Human Decision-Making by Discovering Efficient Strategies for Hierarchical Planning31 Jan 2021 0 repositories listed
-
Deep Reinforcement Learning Aided Monte Carlo Tree Search for MIMO Detection30 Jan 2021 0 repositories listed
-
Deep Reinforcement Learning-Based Product Recommender for Online Advertising30 Jan 2021 0 repositories listed
-
On the Stability of Random Matrix Product with Markovian Noise: Application to Linear Stochastic Approximation and TD Learning30 Jan 2021 0 repositories listed
-
Policy Mirror Descent for Reinforcement Learning: Linear Convergence, New Sampling Complexity, and Generalized Problem Classes30 Jan 2021 0 repositories listed
-
Learning Skills to Navigate without a Master: A Sequential Multi-Policy Reinforcement Learning Algorithm30 Jan 2021 0 repositories listed
-
Reinforcement Learning for Freight Booking Control Problems29 Jan 2021 0 repositories listed
-
Challenges for Using Impact Regularizers to Avoid Negative Side Effects29 Jan 2021 0 repositories listed
-
Learning-based vs Model-free Adaptive Control of a MAV under Wind Gust29 Jan 2021 0 repositories listed
-
Scalable Voltage Control using Structure-Driven Hierarchical Deep Reinforcement Learning29 Jan 2021 0 repositories listed
-
Thermal Control of Laser Powder Bed Fusion Using Deep Reinforcement Learning29 Jan 2021 0 repositories listed
-
CoordiQ : Coordinated Q-learning for Electric Vehicle Charging Recommendation28 Jan 2021 0 repositories listed
-
Reinforcement Learning based Per-antenna Discrete Power Control for Massive MIMO Systems28 Jan 2021 0 repositories listed
-
Universal Trading for Order Execution with Oracle Policy Distillation28 Jan 2021 0 repositories listed
-
Reinforcement Learning Assisted Beamforming for Inter-cell Interference Mitigation in 5G Massive MIMO Networks27 Jan 2021 0 repositories listed
-
Reinforcement Learning for Selective Key Applications in Power Systems: Recent Advances and Future Challenges27 Jan 2021 0 repositories listed
-
Robust Android Malware Detection System against Adversarial Attacks using Q-Learning27 Jan 2021 0 repositories listed
-
Safe Multi-Agent Reinforcement Learning via Shielding27 Jan 2021 0 repositories listed
-
The MineRL 2020 Competition on Sample Efficient Reinforcement Learning using Human Priors26 Jan 2021 0 repositories listed
-
Channel Estimation via Successive Denoising in MIMO OFDM Systems: A Reinforcement Learning Approach25 Jan 2021 0 repositories listed
-
ECOL-R: Encouraging Copying in Novel Object Captioning with Reinforcement Learning25 Jan 2021 0 repositories listed
-
A Methodology for the Development of RL-Based Adaptive Traffic Signal Controllers24 Jan 2021 0 repositories listed
-
Episodic memory governs choices: An RNN-based reinforcement learning model for decision-making task24 Jan 2021 0 repositories listed
-
Fast Sequence Generation with Multi-Agent Reinforcement Learning24 Jan 2021 0 repositories listed
-
GST: Group-Sparse Training for Accelerating Deep Reinforcement Learning24 Jan 2021 0 repositories listed
-
Solving optimal stopping problems with Deep Q-Learning24 Jan 2021 0 repositories listed
-
Feature Selection Using Reinforcement Learning23 Jan 2021 0 repositories listed
-
Decoupled Exploration and Exploitation Policies for Sample-Efficient Reinforcement Learning23 Jan 2021 0 repositories listed
-
Safe Learning and Optimization Techniques: Towards a Survey of the State of the Art23 Jan 2021 0 repositories listed
-
Prior Preference Learning from Experts:Designing a Reward with Active Inference22 Jan 2021 0 repositories listed
-
Adversarial Machine Learning for Flooding Attacks on 5G Radio Access Network Slicing21 Jan 2021 0 repositories listed
-
Model-based Policy Search for Partially Measurable Systems21 Jan 2021 0 repositories listed
-
Flocking and Collision Avoidance for a Dynamic Squad of Fixed-Wing UAVs Using Deep Reinforcement Learning20 Jan 2021 0 repositories listed
-
Deep Reinforcement Learning Optimizes Graphene Nanopores for Efficient Desalination19 Jan 2021 0 repositories listed
-
Dynamic Bicycle Dispatching of Dockless Public Bicycle-sharing Systems using Multi-objective Reinforcement Learning19 Jan 2021 0 repositories listed
-
Meta-Reinforcement Learning for Adaptive Motor Control in Changing Robot Dynamics and Environments19 Jan 2021 0 repositories listed
-
Spatial Assembly: Generative Architecture With Reinforcement Learning, Self Play and Tree Search19 Jan 2021 0 repositories listed
-
18 Jan 2021 0 repositories listed
-
Cooperative and Competitive Biases for Multi-Agent Reinforcement Learning18 Jan 2021 0 repositories listed
-
Deep Reinforcement Learning with Embedded LQR Controllers18 Jan 2021 0 repositories listed
-
Model-Based Reinforcement Learning for Approximate Optimal Control with Temporal Logic Specifications18 Jan 2021 0 repositories listed
-
Regularized Policies are Reward Robust18 Jan 2021 0 repositories listed