Browse State-of-the-Art › Reinforcement Learning › Papers, page 97
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 97 of 132: papers 9,601 to 9,700 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Acceleration of Actor-Critic Deep Reinforcement Learning for Visual Grasping in Clutter by State Representation Learning Based on Disentanglement of a Raw Input Image27 Feb 2020 0 repositories listed
-
Assembly robots with optimized control stiffness through reinforcement learning27 Feb 2020 0 repositories listed
-
Cautious Reinforcement Learning via Distributional Risk in the Dual Domain27 Feb 2020 0 repositories listed
-
Learning in Markov Decision Processes under Constraints27 Feb 2020 0 repositories listed
-
Learning to Resolve Alliance Dilemmas in Many-Player Zero-Sum Games27 Feb 2020 0 repositories listed
-
Opportunities of a Machine Learning-based Decision Support System for Stroke Rehabilitation Assessment27 Feb 2020 0 repositories listed
-
Review, Analysis and Design of a Comprehensive Deep Reinforcement Learning Framework27 Feb 2020 0 repositories listed
-
Sub-Goal Trees -- a Framework for Goal-Based Reinforcement Learning27 Feb 2020 0 repositories listed
-
Towards Modular Algorithm Induction27 Feb 2020 0 repositories listed
-
Cautious Reinforcement Learning with Logical Constraints26 Feb 2020 0 repositories listed
-
Generalized Hindsight for Reinforcement Learning26 Feb 2020 0 repositories listed
-
Learning Navigation Costs from Demonstration in Partially Observable Environments26 Feb 2020 0 repositories listed
-
Neural Ordinary Differential Equation Value Networks for Parametrized Action Spaces26 Feb 2020 0 repositories listed
-
Policy Evaluation Networks26 Feb 2020 0 repositories listed
-
When Do Drivers Concentrate? Attention-based Driver Behavior Modeling With Deep Reinforcement Learning26 Feb 2020 0 repositories listed
-
G-Learner and GIRL: Goal Based Wealth Management with Reinforcement Learning25 Feb 2020 0 repositories listed
-
Metric-Based Imitation Learning Between Two Dissimilar Anthropomorphic Robotic Arms25 Feb 2020 0 repositories listed
-
Model-Based Reinforcement Learning for Physical Systems Without Velocity and Acceleration Measurements25 Feb 2020 0 repositories listed
-
On Reinforcement Learning for Turn-based Zero-sum Markov Games25 Feb 2020 0 repositories listed
-
Scalable Multi-Task Imitation Learning with Autonomous Improvement25 Feb 2020 0 repositories listed
-
Simultaneously Evolving Deep Reinforcement Learning Models using Multifactorial Optimization25 Feb 2020 0 repositories listed
-
A Double Q-Learning Approach for Navigation of Aerial Vehicles with Connectivity Constraint24 Feb 2020 0 repositories listed
-
Backpropamine: training self-modifying neural networks with differentiable neuromodulated plasticity24 Feb 2020 0 repositories listed
-
How Transferable are the Representations Learned by Deep Q Agents?24 Feb 2020 0 repositories listed
-
Millimeter Wave Communications with an Intelligent Reflector: Performance Optimization and Distributional Reinforcement Learning24 Feb 2020 0 repositories listed
-
Q-learning with Uniformly Bounded Variance: Large Discounting is Not a Barrier to Fast Learning24 Feb 2020 0 repositories listed
-
Scalable Multi-Agent Inverse Reinforcement Learning via Actor-Attention-Critic24 Feb 2020 0 repositories listed
-
Computer-inspired Quantum Experiments23 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning with Linear Quadratic Regulator Regions23 Feb 2020 0 repositories listed
-
Near-optimal Regret Bounds for Stochastic Shortest Path23 Feb 2020 0 repositories listed
-
Optimizing Traffic Lights with Multi-agent Deep Reinforcement Learning and V2X communication23 Feb 2020 0 repositories listed
-
Periodic Q-Learning23 Feb 2020 0 repositories listed
-
Rapidly Personalizing Mobile Health Treatment Policies with Limited Data23 Feb 2020 0 repositories listed
-
Wireless 2.0: Towards an Intelligent Radio Environment Empowered by Reconfigurable Meta-Surfaces and Artificial Intelligence23 Feb 2020 0 repositories listed
-
Automatic Data Augmentation via Deep Reinforcement Learning for Effective Kidney Tumor Segmentation22 Feb 2020 0 repositories listed
-
Deep Learning for Ultra-Reliable and Low-Latency Communications in 6G Networks22 Feb 2020 0 repositories listed
-
Guided Constrained Policy Optimization for Dynamic Quadrupedal Robot Locomotion22 Feb 2020 0 repositories listed
-
Vehicle Tracking in Wireless Sensor Networks via Deep Reinforcement Learning22 Feb 2020 0 repositories listed
-
Accelerating Reinforcement Learning with a Directional-Gaussian-Smoothing Evolution Strategy21 Feb 2020 0 repositories listed
-
Data Freshness and Energy-Efficient UAV Navigation Optimization: A Deep Reinforcement Learning Approach21 Feb 2020 0 repositories listed
-
Disentangling Controllable Object through Video Prediction Improves Visual Reinforcement Learning21 Feb 2020 0 repositories listed
-
Minimax-Optimal Off-Policy Evaluation with Linear Function Approximation21 Feb 2020 0 repositories listed
-
On the Search for Feedback in Reinforcement Learning21 Feb 2020 0 repositories listed
-
Risk-Aware Energy Scheduling for Edge Computing with Microgrid: A Multi-Agent Deep Reinforcement Learning Approach21 Feb 2020 0 repositories listed
-
Adaptive Temporal Difference Learning with Linear Function Approximation20 Feb 2020 0 repositories listed
-
20 Feb 2020 0 repositories listed
-
Enhanced Adversarial Strategically-Timed Attacks against Deep Reinforcement Learning20 Feb 2020 0 repositories listed
-
Multi-Agent Meta-Reinforcement Learning for Self-Powered and Sustainable Edge Computing Systems20 Feb 2020 0 repositories listed
-
Multi-Agent Reinforcement Learning as a Computational Tool for Language Evolution Research: Historical Context and Future Challenges20 Feb 2020 0 repositories listed
-
oIRL: Robust Adversarial Inverse Reinforcement Learning with Temporally Extended Actions20 Feb 2020 0 repositories listed
-
Debiased Off-Policy Evaluation for Recommendation Systems20 Feb 2020 0 repositories listed
-
Curriculum in Gradient-Based Meta-Reinforcement Learning19 Feb 2020 0 repositories listed
-
Keep Doing What Worked: Behavioral Modelling Priors for Offline Reinforcement Learning19 Feb 2020 0 repositories listed
-
Neural Architecture Search For Fault Diagnosis19 Feb 2020 0 repositories listed
-
Optimistic Policy Optimization with Bandit Feedback19 Feb 2020 0 repositories listed
-
UAV Aided Search and Rescue Operation Using Reinforcement Learning19 Feb 2020 0 repositories listed
-
Using AI for Mitigating the Impact of Network Delay in Cloud-based Intelligent Traffic Signal Control19 Feb 2020 0 repositories listed
-
Value-driven Hindsight Modelling19 Feb 2020 0 repositories listed
-
Empirical Policy Evaluation with Supergraphs18 Feb 2020 0 repositories listed
-
KoGuN: Accelerating Deep Reinforcement Learning via Integrating Human Suboptimal Knowledge18 Feb 2020 0 repositories listed
-
MoTiAC: Multi-Objective Actor-Critics for Real-Time Bidding18 Feb 2020 0 repositories listed
-
Multi-Issue Bargaining With Deep Reinforcement Learning18 Feb 2020 0 repositories listed
-
Adaptive Experience Selection for Policy Gradient17 Feb 2020 0 repositories listed
-
GACEM: Generalized Autoregressive Cross Entropy Method for Multi-Modal Black Box Constraint Satisfaction17 Feb 2020 0 repositories listed
-
Learning Zero-Sum Simultaneous-Move Markov Games Using Function Approximation and Correlated Equilibrium17 Feb 2020 0 repositories listed
-
Reinforcement learning for the privacy preservation and manipulation of eye tracking data17 Feb 2020 0 repositories listed
-
Reward Design for Driver Repositioning Using Multi-Agent Reinforcement Learning17 Feb 2020 0 repositories listed
-
Investigating Simple Object Representations in Model-Free Deep Reinforcement Learning16 Feb 2020 0 repositories listed
-
TempLe: Learning Template of Transitions for Sample Efficient Multi-task RL16 Feb 2020 0 repositories listed
-
MRRC: Multiple Role Representation Crossover Interpretation for Image Captioning With R-CNN Feature Distribution Composition (FDC)15 Feb 2020 0 repositories listed
-
Non-asymptotic Convergence of Adam-type Reinforcement Learning Algorithms under Markovian Sampling15 Feb 2020 0 repositories listed
-
The Archimedean trap: Why traditional reinforcement learning will probably not yield AGI15 Feb 2020 0 repositories listed
-
Frequency-based Search-control in Dyna14 Feb 2020 0 repositories listed
-
Learning Functionally Decomposed Hierarchies for Continuous Control Tasks with Path Planning14 Feb 2020 0 repositories listed
-
Resource Management in Wireless Networks via Multi-Agent Deep Reinforcement Learning14 Feb 2020 0 repositories listed
-
Stable Training of DNN for Speech Enhancement based on Perceptually-Motivated Black-Box Cost Function14 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning-Based Beam Tracking for Low-Latency Services in Vehicular Networks13 Feb 2020 0 repositories listed
-
Fast Reinforcement Learning for Anti-jamming Communications13 Feb 2020 0 repositories listed
-
Improving Generalization of Reinforcement Learning with Minimax Distributional Soft Actor-Critic13 Feb 2020 0 repositories listed
-
MODRL/D-AM: Multiobjective Deep Reinforcement Learning Algorithm Using Decomposition and Attention Model for Multiobjective Optimization13 Feb 2020 0 repositories listed
-
Multi-Vehicle Routing Problems with Soft Time Windows: A Multi-Agent Reinforcement Learning Approach13 Feb 2020 0 repositories listed
-
On the Sensory Commutativity of Action Sequences for Embodied Agents13 Feb 2020 0 repositories listed
-
XCS Classifier System with Experience Replay13 Feb 2020 0 repositories listed
-
A Tensor Network Approach to Finite Markov Decision Processes12 Feb 2020 0 repositories listed
-
Learning Multi-Agent Coordination through Connectivity-driven Communication12 Feb 2020 0 repositories listed
-
Data Efficient Training for Reinforcement Learning with Adaptive Behavior Policy Sharing12 Feb 2020 0 repositories listed
-
Intrinsic Motivation for Encouraging Synergistic Behavior12 Feb 2020 0 repositories listed
-
Regret Bounds for Discounted MDPs12 Feb 2020 0 repositories listed
-
AI Online Filters to Real World Image Recognition11 Feb 2020 0 repositories listed
-
Best of Both Worlds: AutoML Codesign of a CNN and its Hardware Accelerator11 Feb 2020 0 repositories listed
-
Confounding-Robust Policy Evaluation in Infinite-Horizon Reinforcement Learning11 Feb 2020 0 repositories listed
-
HMRL: Hyper-Meta Learning for Sparse Reward Reinforcement Learning Problem11 Feb 2020 0 repositories listed
-
Learning Structured Communication for Multi-agent Reinforcement Learning11 Feb 2020 0 repositories listed
-
Learning to Switch Among Agents in a Team via 2-Layer Markov Decision Processes11 Feb 2020 0 repositories listed
-
Machine Learning Approaches For Motor Learning: A Short Review11 Feb 2020 0 repositories listed
-
Reinforcement Learning for POMDP: Partitioned Rollout and Policy Iteration with Application to Autonomous Sequential Repair Problems11 Feb 2020 0 repositories listed
-
Towards Intelligent Pick and Place Assembly of Individualized Products Using Reinforcement Learning11 Feb 2020 0 repositories listed
-
Convergence Guarantees of Policy Optimization Methods for Markovian Jump Linear Systems10 Feb 2020 0 repositories listed
-
Interpretable Off-Policy Evaluation in Reinforcement Learning by Highlighting Influential Transitions10 Feb 2020 0 repositories listed
-
On Reward Shaping for Mobile Robot Navigation: A Reinforcement Learning and SLAM Based Approach10 Feb 2020 0 repositories listed