Browse State-of-the-Art › Reinforcement Learning › Papers, page 96
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 96 of 132: papers 9,501 to 9,600 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Multi-Agent Reinforcement Learning for Problems with Combined Individual and Team Reward24 Mar 2020 0 repositories listed
-
Q-Learning in Regularized Mean-field Games24 Mar 2020 0 repositories listed
-
Importance of using appropriate baselines for evaluation of data-efficiency in deep reinforcement learning for Atari23 Mar 2020 0 repositories listed
-
Incorporating Relational Background Knowledge into Reinforcement Learning via Differentiable Inductive Logic Programming23 Mar 2020 0 repositories listed
-
Learning to Walk: Spike Based Reinforcement Learning for Hexapod Robot Central Pattern Generation22 Mar 2020 0 repositories listed
-
Reinforcement Learning in Economics and Finance22 Mar 2020 0 repositories listed
-
Accelerating Deep Reinforcement Learning With the Aid of Partial Model: Energy-Efficient Predictive Video Streaming21 Mar 2020 0 repositories listed
-
Autonomous UAV Navigation: A DDPG-based Deep Reinforcement Learning Approach21 Mar 2020 0 repositories listed
-
Comprehensive Review of Deep Reinforcement Learning Methods and Applications in Economics21 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning with Robust and Smooth Policy21 Mar 2020 0 repositories listed
-
Distributed Reinforcement Learning for Cooperative Multi-Robot Object Manipulation21 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning with Weighted Q-Learning20 Mar 2020 0 repositories listed
-
Deep Sets for Generalization in RL20 Mar 2020 0 repositories listed
-
Deep Constrained Q-learning20 Mar 2020 0 repositories listed
-
Exchangeable Input Representations for Reinforcement Learning19 Mar 2020 0 repositories listed
-
Reinforcement learning enabled cooperative spectrum sensing in cognitive radio networks19 Mar 2020 0 repositories listed
-
Towards Cognitive Routing based on Deep Reinforcement Learning19 Mar 2020 0 repositories listed
-
Adaptive-Step Graph Meta-Learner for Few-Shot Graph Classification18 Mar 2020 0 repositories listed
-
Generating Socially Acceptable Perturbations for Efficient Evaluation of Autonomous Vehicles18 Mar 2020 0 repositories listed
-
Placement Optimization with Deep Reinforcement Learning18 Mar 2020 0 repositories listed
-
Viewport-Aware Deep Reinforcement Learning Approach for 360ᵒ Video Caching18 Mar 2020 0 repositories listed
-
Stop-and-Go: Exploring Backdoor Attacks on Deep Reinforcement Learning-based Traffic Congestion Control Systems17 Mar 2020 0 repositories listed
-
Improving Performance in Reinforcement Learning by Breaking Generalization in Neural Networks16 Mar 2020 0 repositories listed
-
Reinforcement Learning for Electricity Network Operation16 Mar 2020 0 repositories listed
-
Value Variance Minimization for Learning Approximate Equilibrium in Aggregation Systems16 Mar 2020 0 repositories listed
-
Active Perception and Representation for Robotic Manipulation15 Mar 2020 0 repositories listed
-
Model-based Reinforcement Learning for Decentralized Multiagent Rendezvous15 Mar 2020 0 repositories listed
-
A General Framework for Learning Mean-Field Games13 Mar 2020 0 repositories listed
-
Application of Deep Q-Network in Portfolio Management13 Mar 2020 0 repositories listed
-
Optimizing Medical Treatment for Sepsis in Intensive Care: from Reinforcement Learning to Pre-Trial Evaluation13 Mar 2020 0 repositories listed
-
13 Mar 2020 0 repositories listed
-
Analysis of Hyper-Parameters for Small Games: Iterations or Epochs in Self-Play?12 Mar 2020 0 repositories listed
-
Analyzing Visual Representations in Embodied Navigation Tasks12 Mar 2020 0 repositories listed
-
Heterogeneous Relational Reasoning in Knowledge Graphs with Reinforcement Learning12 Mar 2020 0 repositories listed
-
Adaptive Control and Regret Minimization in Linear Quadratic Gaussian (LQG) Setting12 Mar 2020 0 repositories listed
-
Automatic Curriculum Learning For Deep RL: A Short Survey10 Mar 2020 0 repositories listed
-
Curriculum Learning for Reinforcement Learning Domains: A Framework and Survey10 Mar 2020 0 repositories listed
-
Indirect and Direct Training of Spiking Neural Networks for End-to-End Control of a Lane-Keeping Vehicle10 Mar 2020 0 repositories listed
-
Learning to be Global Optimizer10 Mar 2020 0 repositories listed
-
Privacy-Cost Management in Smart Meters Using Deep Reinforcement Learning10 Mar 2020 0 repositories listed
-
Reinforcement Learning for Mitigating Intermittent Interference in Terahertz Communication Networks10 Mar 2020 0 repositories listed
-
SQUIRL: Robust and Efficient Learning from Video Demonstration of Long-Horizon Robotic Manipulation Tasks10 Mar 2020 0 repositories listed
-
10 Mar 2020 0 repositories listed Syntology 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 7 harvested samples) · 3 pointer-only (licence)
-
Advancing Renewable Electricity Consumption With Reinforcement Learning9 Mar 2020 0 repositories listed
-
A Multi-Agent Reinforcement Learning Approach For Safe and Efficient Behavior Planning Of Connected Autonomous Vehicles9 Mar 2020 0 repositories listed
-
Efficiency and Equity are Both Essential: A Generalized Traffic Signal Controller with Deep Reinforcement Learning9 Mar 2020 0 repositories listed
-
Human AI interaction loop training: New approach for interactive reinforcement learning9 Mar 2020 0 repositories listed
-
Q* Approximation Schemes for Batch Reinforcement Learning: A Theoretical Comparison9 Mar 2020 0 repositories listed
-
Software-Level Accuracy Using Stochastic Computing With Charge-Trap-Flash Based Weight Matrix9 Mar 2020 0 repositories listed
-
Transfer Reinforcement Learning under Unobserved Contextual Information9 Mar 2020 0 repositories listed
-
Zooming for Efficient Model-Free Reinforcement Learning in Metric Spaces9 Mar 2020 0 repositories listed
-
Deep Adversarial Reinforcement Learning for Object Disentangling8 Mar 2020 0 repositories listed
-
Generative Adversarial Imitation Learning with Neural Networks: Global Optimality and Convergence Rate8 Mar 2020 0 repositories listed
-
Reinforcement Learning Based Cooperative Coded Caching under Dynamic Popularities in Ultra-Dense Networks8 Mar 2020 0 repositories listed
-
Convergence of Q-value in case of Gaussian rewards7 Mar 2020 0 repositories listed
-
Reinforcement Learning for Combinatorial Optimization: A Survey7 Mar 2020 0 repositories listed
-
Cost-Sensitive Portfolio Selection via Deep Reinforcement Learning6 Mar 2020 0 repositories listed
-
Lane-Merging Using Policy-based Reinforcement Learning and Post-Optimization6 Mar 2020 0 repositories listed
-
Smart Train Operation Algorithms based on Expert Knowledge and Reinforcement Learning6 Mar 2020 0 repositories listed
-
A Geometric Perspective on Visual Imitation Learning5 Mar 2020 0 repositories listed
-
Balance Between Efficient and Effective Learning: Dense2Sparse Reward Shaping for Robot Manipulation with Environment Uncertainty5 Mar 2020 0 repositories listed
-
BERT as a Teacher: Contextual Embeddings for Sequence-Level Reward5 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning-BasedRobust Protection in DER-Rich Distribution Grids5 Mar 2020 0 repositories listed
-
Distributional Robustness and Regularization in Reinforcement Learning5 Mar 2020 0 repositories listed
-
Efficient and Effective Similar Subtrajectory Search with Deep Reinforcement Learning5 Mar 2020 0 repositories listed
-
Reward Design in Cooperative Multi-agent Reinforcement Learning for Packet Routing5 Mar 2020 0 repositories listed
-
Dynamic Experience Replay4 Mar 2020 0 repositories listed
-
Efficient statistical validation with edge cases to evaluate Highly Automated Vehicles4 Mar 2020 0 repositories listed
-
Neural-Network Heuristics for Adaptive Bayesian Quantum Estimation4 Mar 2020 0 repositories listed
-
On Hyper-parameter Tuning for Stochastic Optimization Algorithms4 Mar 2020 0 repositories listed
-
Privacy-Aware Time-Series Data Sharing with Deep Reinforcement Learning4 Mar 2020 0 repositories listed
-
Deep Reinforcement Learning for QoS-Constrained Resource Allocation in Multiservice Networks3 Mar 2020 0 repositories listed
-
Efficient Exploration in Constrained Environments with Goal-Oriented Reference Path3 Mar 2020 0 repositories listed
-
Learning Context-aware Task Reasoning for Efficient Meta-reinforcement Learning3 Mar 2020 0 repositories listed
-
Safe Reinforcement Learning for Autonomous Vehicles through Parallel Constrained Policy Optimization3 Mar 2020 0 repositories listed
-
Relevance-Guided Modeling of Object Dynamics for Reinforcement Learning3 Mar 2020 0 repositories listed
-
Adaptive Structural Hyper-Parameter Configuration by Q-Learning2 Mar 2020 0 repositories listed
-
Cluster-Based Social Reinforcement Learning2 Mar 2020 0 repositories listed
-
Formal Controller Synthesis for Continuous-Space MDPs via Model-Free Reinforcement Learning2 Mar 2020 0 repositories listed
-
Gaussian Process Policy Optimization2 Mar 2020 0 repositories listed
-
Learning Force Control for Contact-rich Manipulation Tasks with Rigid Position-controlled Robots2 Mar 2020 0 repositories listed
-
Real-World Human-Robot Collaborative Reinforcement Learning2 Mar 2020 0 repositories listed
-
Risk-Averse Learning by Temporal Difference Methods2 Mar 2020 0 repositories listed
-
Scaling Up Multiagent Reinforcement Learning for Robotic Systems: Learn an Adaptive Sparse Communication Graph2 Mar 2020 0 repositories listed
-
Upper Confidence Primal-Dual Reinforcement Learning for CMDP with Adversarial Loss2 Mar 2020 0 repositories listed
-
Dynamic Queue-Jump Lane for Emergency Vehicles under Partially Connected Settings: A Multi-Agent Deep Reinforcement Learning Approach2 Mar 2020 0 repositories listed
-
Fully Asynchronous Policy Evaluation in Distributed Reinforcement Learning over Networks1 Mar 2020 0 repositories listed
-
Learn Task First or Learn Human Partner First: A Hierarchical Task Decomposition Method for Human-Robot Cooperation1 Mar 2020 0 repositories listed
-
PlaNet of the Bayesians: Reconsidering and Improving Deep Planning Network by Incorporating Bayesian Inference1 Mar 2020 0 repositories listed
-
Provably Efficient Safe Exploration via Primal-Dual Policy Optimization1 Mar 2020 0 repositories listed
-
Contextual Policy Transfer in Reinforcement Learning Domains via Deep Mixtures-of-Experts29 Feb 2020 0 repositories listed
-
Learning Near Optimal Policies with Low Inherent Bellman Error29 Feb 2020 0 repositories listed
-
Do optimization methods in deep learning applications matter?28 Feb 2020 0 repositories listed
-
Efficiently Guiding Imitation Learning Agents with Human Gaze28 Feb 2020 0 repositories listed
-
Jointly Learning to Recommend and Advertise28 Feb 2020 0 repositories listed
-
Mixed Reinforcement Learning with Additive Stochastic Uncertainty28 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning for FlipIt Security Game28 Feb 2020 0 repositories listed
-
Reinforcement Learning through Active Inference28 Feb 2020 0 repositories listed
-
A Self-Tuning Actor-Critic Algorithm28 Feb 2020 0 repositories listed
-
A Visual Communication Map for Multi-Agent Deep Reinforcement Learning27 Feb 2020 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.