Browse State-of-the-Art › Reinforcement Learning › Papers, page 119
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 119 of 132: papers 11,801 to 11,900 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Using Reinforcement Learning with Partial Vehicle Detection for Intelligent Traffic Signal Control4 Jul 2018 0 repositories listed
-
Region Growing Curriculum Generation for Reinforcement Learning4 Jul 2018 0 repositories listed
-
Supervised Reinforcement Learning with Recurrent Neural Network for Dynamic Treatment Recommendation4 Jul 2018 0 repositories listed
-
Transfer with Model Features in Reinforcement Learning4 Jul 2018 0 repositories listed
-
Human-level performance in first-person multiplayer games with population-based deep reinforcement learning3 Jul 2018 0 repositories listed
-
Speeding up the Metabolism in E-commerce by Reinforcement Mechanism Design2 Jul 2018 0 repositories listed
-
1 Jul 2018 0 repositories listed
-
A Reinforcement Learning Neural Network for Robotic Manipulator Control1 Jul 2018 0 repositories listed
-
Beyond the One-Step Greedy Approach in Reinforcement Learning1 Jul 2018 0 repositories listed
-
Beyond Winning and Losing: Modeling Human Motivations and Behaviors Using Inverse Reinforcement Learning1 Jul 2018 0 repositories listed
-
Decoupling Gradient-Like Learning Rules from Representations1 Jul 2018 0 repositories listed
-
Deep Reinforcement Learning for NLP1 Jul 2018 0 repositories listed
-
Feudal Dialogue Management with Jointly Learned Feature Extractors1 Jul 2018 0 repositories listed
-
Improved Regret Bounds for Thompson Sampling in Linear Quadratic Control Problems1 Jul 2018 0 repositories listed
-
Learning Hierarchical Structures On-The-Fly with a Recurrent-Recursive Model for Sequences1 Jul 2018 0 repositories listed
-
Learning to Act in Decentralized Partially Observable MDPs1 Jul 2018 0 repositories listed
-
Learning to Coordinate with Coordination Graphs in Repeated Single-Stage Multi-Agent Decision Problems1 Jul 2018 0 repositories listed
-
Learning to Explore via Meta-Policy Gradient1 Jul 2018 0 repositories listed
-
Mix & Match - Agent Curricula for Reinforcement Learning1 Jul 2018 0 repositories listed
-
Neural User Simulation for Corpus-based Policy Optimisation of Spoken Dialogue Systems1 Jul 2018 0 repositories listed
-
Policy and Value Transfer in Lifelong Reinforcement Learning1 Jul 2018 0 repositories listed
-
Policy Optimization with Demonstrations1 Jul 2018 0 repositories listed
-
Scalable Bilinear Pi Learning Using State and Action Features1 Jul 2018 0 repositories listed
-
Spotlight: Optimizing Device Placement for Training Deep Neural Networks1 Jul 2018 0 repositories listed
-
State Abstractions for Lifelong Reinforcement Learning1 Jul 2018 0 repositories listed
-
The Importance of Recommender and Feedback Features in a Pronunciation Learning Aid1 Jul 2018 0 repositories listed
-
Towards Mixed Optimization for Reinforcement Learning with Program Synthesis1 Jul 2018 0 repositories listed
-
Understanding and Simplifying One-Shot Architecture Search1 Jul 2018 0 repositories listed
-
Hierarchical Reinforcement Learning with Abductive Planning28 Jun 2018 0 repositories listed
-
MONAS: Multi-Objective Neural Architecture Search using Reinforcement Learning27 Jun 2018 0 repositories listed
-
Adversarial Active Exploration for Inverse Dynamics Model Learning26 Jun 2018 0 repositories listed
-
Deep Generative Models with Learnable Knowledge Constraints26 Jun 2018 0 repositories listed
-
Learning Existing Social Conventions via Observationally Augmented Self-Play26 Jun 2018 0 repositories listed
-
Multi-agent Inverse Reinforcement Learning for Certain General-sum Stochastic Games26 Jun 2018 0 repositories listed
-
Deep Reinforcement Learning: An Overview23 Jun 2018 0 repositories listed
-
Human-Interactive Subgoal Supervision for Efficient Inverse Reinforcement Learning22 Jun 2018 0 repositories listed
-
Learning-to-Ask: Knowledge Acquisition via 20 Questions22 Jun 2018 0 repositories listed
-
Many-Goals Reinforcement Learning22 Jun 2018 0 repositories listed
-
A New Approach for Resource Scheduling with Deep Reinforcement Learning21 Jun 2018 0 repositories listed
-
Expanding the Active Inference Landscape: More Intrinsic Motivations in the Perception-Action Loop21 Jun 2018 0 repositories listed
-
A Dissection of Overfitting and Generalization in Continuous Reinforcement Learning20 Jun 2018 0 repositories listed
-
Learning Neural Parsers with Deterministic Differentiable Imitation Learning20 Jun 2018 0 repositories listed
-
Reinforcement Learning using Augmented Neural Networks20 Jun 2018 0 repositories listed
-
Skilled Experience Catalogue: A Skill-Balancing Mechanism for Non-Player Characters using Reinforcement Learning20 Jun 2018 0 repositories listed
-
A Survey of Inverse Reinforcement Learning: Challenges, Methods and Progress18 Jun 2018 0 repositories listed
-
A unified strategy for implementing curiosity and empowerment driven reinforcement learning18 Jun 2018 0 repositories listed
-
Learning from Outside the Viability Kernel: Why we Should Build Robots that can Fall with Grace18 Jun 2018 0 repositories listed
-
Learning Policy Representations in Multiagent Systems17 Jun 2018 0 repositories listed
-
Handling Cold-Start Collaborative Filtering with Reinforcement Learning16 Jun 2018 0 repositories listed
-
An Online Prediction Algorithm for Reinforcement Learning with Linear Function Approximation using Cross Entropy Method15 Jun 2018 0 repositories listed
-
Improving width-based planning with compact policies15 Jun 2018 0 repositories listed
-
Multi-Level Policy and Reward Reinforcement Learning for Image Captioning15 Jun 2018 0 repositories listed
-
Adaptive Shooting for Bots in First Person Shooter Games Using Reinforcement Learning14 Jun 2018 0 repositories listed
-
Deep Reinforcement Learning for Dynamic Urban Transportation Problems14 Jun 2018 0 repositories listed
-
Qualitative Measurements of Policy Discrepancy for Return-Based Deep Q-Network14 Jun 2018 0 repositories listed
-
Automatic formation of the structure of abstract machines in hierarchical reinforcement learning with state clustering13 Jun 2018 0 repositories listed
-
Learning to Shoot in First Person Shooter Games by Stabilizing Actions and Clustering Rewards for Reinforcement Learning13 Jun 2018 0 repositories listed
-
Reinforcement Learning with Function-Valued Action Spaces for Partial Differential Equation Control13 Jun 2018 0 repositories listed
-
Accelerating Imitation Learning with Predictive Models12 Jun 2018 0 repositories listed
-
Improving Regression Performance with Distributional Losses12 Jun 2018 0 repositories listed
-
Meta-Learning Transferable Active Learning Policies by Deep Reinforcement Learning12 Jun 2018 0 repositories listed
-
Multi-Agent Deep Reinforcement Learning with Human Strategies12 Jun 2018 0 repositories listed
-
Unsupervised Meta-Learning for Reinforcement Learning12 Jun 2018 0 repositories listed
-
An Efficient, Generalized Bellman Update For Cooperative Inverse Reinforcement Learning11 Jun 2018 0 repositories listed
-
Context-Aware Policy Reuse11 Jun 2018 0 repositories listed
-
Deep Curiosity Loops in Social Environments10 Jun 2018 0 repositories listed
-
Implicit Policy for Reinforcement Learning10 Jun 2018 0 repositories listed
-
Automatic View Planning with Multi-scale Deep Reinforcement Learning Agents8 Jun 2018 0 repositories listed
-
Continuous-time Value Function Approximation in Reproducing Kernel Hilbert Spaces8 Jun 2018 0 repositories listed
-
Fidelity-based Probabilistic Q-learning for Control of Quantum Systems8 Jun 2018 0 repositories listed
-
Program Synthesis Through Reinforcement Learning Guided Tree Search8 Jun 2018 0 repositories listed
-
Self-Consistent Trajectory Autoencoder: Hierarchical Reinforcement Learning with Trajectory Embeddings7 Jun 2018 0 repositories listed
-
Simplifying Reward Design through Divide-and-Conquer7 Jun 2018 0 repositories listed
-
A Finite Time Analysis of Temporal Difference Learning With Linear Function Approximation6 Jun 2018 0 repositories listed
-
Meta-Learning by the Baldwin Effect6 Jun 2018 0 repositories listed
-
Discovering and Removing Exogenous State Variables and Rewards for Reinforcement Learning5 Jun 2018 0 repositories listed
-
Mix&Match - Agent Curricula for Reinforcement Learning5 Jun 2018 0 repositories listed
-
The Effect of Planning Shape on Dyna-style Planning in High-dimensional State Spaces5 Jun 2018 0 repositories listed
-
Adversarial Reinforcement Learning Framework for Benchmarking Collision Avoidance Mechanisms in Autonomous Vehicles4 Jun 2018 0 repositories listed
-
Penalizing side effects using stepwise relative reachability4 Jun 2018 0 repositories listed
-
Mitigation of Policy Manipulation Attacks on Deep Q-Networks with Parameter-Space Noise4 Jun 2018 0 repositories listed
-
Relational inductive bias for physical construction in humans and machines4 Jun 2018 0 repositories listed
-
Sequential Test for the Lowest Mean: From Thompson to Murphy Sampling4 Jun 2018 0 repositories listed
-
Building Advanced Dialogue Managers for Goal-Oriented Dialogue Systems3 Jun 2018 0 repositories listed
-
Exploration in Structured Reinforcement Learning3 Jun 2018 0 repositories listed
-
Multi-Agent Reinforcement Learning via Double Averaging Primal-Dual Optimization3 Jun 2018 0 repositories listed
-
DAQN: Deep Auto-encoder and Q-Network2 Jun 2018 0 repositories listed
-
Deep Pepper: Expert Iteration based Chess agent in the Reinforcement Learning Setting2 Jun 2018 0 repositories listed
-
Efficient Entropy for Policy Gradient with Multidimensional Action Space2 Jun 2018 0 repositories listed
-
Internal Model from Observations for Reward Shaping2 Jun 2018 0 repositories listed
-
A Reinforcement Learning Approach to Age of Information in Multi-User Networks1 Jun 2018 0 repositories listed
-
Bootstrapping a Neural Conversational Agent with Dialogue Self-Play, Crowdsourcing and On-Line Reinforcement Learning1 Jun 2018 0 repositories listed
-
Deep Curiosity Search: Intra-Life Exploration Can Improve Performance on Challenging Deep Reinforcement Learning Problems1 Jun 2018 0 repositories listed
-
1 Jun 2018 0 repositories listed
-
Egocentric Activity Recognition on a Budget1 Jun 2018 0 repositories listed
-
End-to-End Learning of Task-Oriented Dialogs1 Jun 2018 0 repositories listed
-
Environment Upgrade Reinforcement Learning for Non-Differentiable Multi-Stage Pipelines1 Jun 2018 0 repositories listed
-
Equivalence Between Wasserstein and Value-Aware Loss for Model-based Reinforcement Learning1 Jun 2018 0 repositories listed
-
Fast Exploration with Simplified Models and Approximately Optimistic Planning in Model Based Reinforcement Learning1 Jun 2018 0 repositories listed
-
GraphBit: Bitwise Interaction Mining via Deep Reinforcement Learning1 Jun 2018 0 repositories listed