Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 148
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 148 of 152: papers 14,701 to 14,800 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Show, Attend and Interact: Perceivable Human-Robot Social Interaction through Neural Attention Q-Network28 Feb 2017 0 repositories listed
-
27 Feb 2017 0 repositories listed
-
Learning Control for Air Hockey Striking using Deep Reinforcement Learning26 Feb 2017 0 repositories listed
-
Stochastic Variance Reduction Methods for Policy Evaluation25 Feb 2017 0 repositories listed
-
Changing Model Behavior at Test-Time Using Reinforcement Learning24 Feb 2017 0 repositories listed
-
Control of Gene Regulatory Networks with Noisy Measurements and Uncertain Inputs24 Feb 2017 0 repositories listed
-
Online Meta-learning by Parallel Algorithm Competition24 Feb 2017 0 repositories listed
-
Robot gains Social Intelligence through Multimodal Deep Reinforcement Learning24 Feb 2017 0 repositories listed
-
Automatic Representation for Lifetime Value Recommender Systems23 Feb 2017 0 repositories listed
-
Data Distillation for Controlling Specificity in Dialogue Generation22 Feb 2017 0 repositories listed
-
Reinforcement Learning Based Argument Component Detection21 Feb 2017 0 repositories listed
-
Learning to Repeat: Fine Grained Action Repetition for Deep Reinforcement Learning20 Feb 2017 0 repositories listed
-
Collaborative Deep Reinforcement Learning for Joint Object Search18 Feb 2017 0 repositories listed
-
Batch Policy Gradient Methods for Improving Neural Conversation Models10 Feb 2017 0 repositories listed
-
10 Feb 2017 0 repositories listed
-
Semi-Supervised QA with Generative Domain-Adaptive Nets7 Feb 2017 0 repositories listed
-
Uncertainty-Aware Reinforcement Learning for Collision Avoidance3 Feb 2017 0 repositories listed
-
Deep Reinforcement Learning for Robotic Manipulation-The state of the art31 Jan 2017 0 repositories listed
-
Deep Reinforcement Learning for Visual Object Tracking in Videos31 Jan 2017 0 repositories listed
-
Expert Level control of Ramp Metering based on Multi-task Deep Reinforcement Learning30 Jan 2017 0 repositories listed
-
Flow Navigation by Smart Microswimmers via Reinforcement Learning30 Jan 2017 0 repositories listed
-
Reinforcement Learning Algorithm Selection30 Jan 2017 0 repositories listed
-
Artificial Intelligence Approaches To UCAV Autonomy24 Jan 2017 0 repositories listed
-
Binary Matrix Guessing Problem22 Jan 2017 0 repositories listed
-
Basic protocols in quantum reinforcement learning with superconducting circuits18 Jan 2017 0 repositories listed
-
Agent-Agnostic Human-in-the-Loop Reinforcement Learning15 Jan 2017 0 repositories listed
-
Scalable and Incremental Learning of Gaussian Mixture Models14 Jan 2017 0 repositories listed
-
Reinforcement Learning based Embodied Agents Modelling Human Users Through Interaction and Multi-Sensory Perception9 Jan 2017 0 repositories listed
-
A Review of Neural Network Based Machine Learning Approaches for Rotor Angle Stability Control5 Jan 2017 0 repositories listed
-
Toward negotiable reinforcement learning: shifting priorities in Pareto optimal sequential decision-making5 Jan 2017 0 repositories listed
-
First-Person Activity Forecasting with Online Inverse Reinforcement Learning22 Dec 2016 0 repositories listed
-
Non-Deterministic Policy Improvement Stabilizes Approximated Reinforcement Learning22 Dec 2016 0 repositories listed
-
On the function approximation error for risk-sensitive reinforcement learning22 Dec 2016 0 repositories listed
-
Loss is its own Reward: Self-Supervision for Reinforcement Learning21 Dec 2016 0 repositories listed
-
Unsupervised Perceptual Rewards for Imitation Learning20 Dec 2016 0 repositories listed
-
Sample-efficient Deep Reinforcement Learning for Dialog Control18 Dec 2016 0 repositories listed
-
Learning to predict where to look in interactive environments using deep recurrent q-learning17 Dec 2016 0 repositories listed
-
Reinforcement Learning Using Quantum Boltzmann Machines17 Dec 2016 0 repositories listed
-
Deep Reinforcement Learning with Successor Features for Navigation across Similar Environments16 Dec 2016 0 repositories listed
-
Separation of Concerns in Reinforcement Learning15 Dec 2016 0 repositories listed
-
End-to-End Deep Reinforcement Learning for Lane Keeping Assist13 Dec 2016 0 repositories listed
-
Incorporating Human Domain Knowledge into Large Scale Cost Function Learning13 Dec 2016 0 repositories listed
-
Response to Comment on 'Perceptual Learning Incepted by Decoded fMRI Neurofeedback Without Stimulus Presentation'; How can a decoded neurofeedback method (DecNef) lead to successful reinforcement and visual perceptual learning?13 Dec 2016 0 repositories listed
-
Learning to Drive using Inverse Reinforcement Learning and Deep Q-Networks12 Dec 2016 0 repositories listed
-
Online Reinforcement Learning for Real-Time Exploration in Continuous State and Action Markov Decision Processes12 Dec 2016 0 repositories listed
-
PoseAgent: Budget-Constrained 6D Object Pose Estimation via Reinforcement Learning12 Dec 2016 0 repositories listed
-
Reinforcement Learning With Temporal Logic Rewards11 Dec 2016 0 repositories listed
-
Towards deep learning with spiking neurons in energy based models with contrastive Hebbian plasticity9 Dec 2016 0 repositories listed
-
Hierarchy through Composition with Linearly Solvable Markov Decision Processes8 Dec 2016 0 repositories listed
-
Stochastic Primal-Dual Methods and Sample Complexity of Reinforcement Learning8 Dec 2016 0 repositories listed
-
Towards Information-Seeking Agents8 Dec 2016 0 repositories listed
-
Deep Learning of Robotic Tasks without a Simulator using Strong and Weak Human Supervision4 Dec 2016 0 repositories listed
-
Learning to superoptimize programs - Workshop Version4 Dec 2016 0 repositories listed
-
Adaptive optimal training of animal behavior1 Dec 2016 0 repositories listed
-
Bootstrapping incremental dialogue systems: using linguistic knowledge to learn from minimal data1 Dec 2016 0 repositories listed
-
Generalizing Skills with Semi-Supervised Reinforcement Learning1 Dec 2016 0 repositories listed
-
Linear Feature Encoding for Reinforcement Learning1 Dec 2016 0 repositories listed
-
Showing versus doing: Teaching by demonstration1 Dec 2016 0 repositories listed
-
Exploration for Multi-task Reinforcement Learning with Deep Generative Models29 Nov 2016 0 repositories listed
-
Improving Policy Gradient by Exploring Under-appreciated Rewards28 Nov 2016 0 repositories listed
-
Learning to Compose Words into Sentences with Reinforcement Learning28 Nov 2016 0 repositories listed
-
Nonparametric General Reinforcement Learning28 Nov 2016 0 repositories listed
-
Multiscale Inverse Reinforcement Learning using Diffusion Wavelets24 Nov 2016 0 repositories listed
-
Recurrent Attention Models for Depth-Based Person Identification22 Nov 2016 0 repositories listed
-
A Deep Learning Approach for Joint Video Frame and Reward Prediction in Atari Games21 Nov 2016 0 repositories listed
-
Memory Lens: How Much Memory Does an Agent Use?21 Nov 2016 0 repositories listed
-
Options Discovery with Budgeted Reinforcement Learning21 Nov 2016 0 repositories listed
-
Reinforcement Learning in Rich-Observation MDPs using Spectral Methods11 Nov 2016 0 repositories listed
-
Fairness in Reinforcement Learning9 Nov 2016 0 repositories listed
-
Sequence Tutor: Conservative Fine-Tuning of Sequence Generation Models with KL-control9 Nov 2016 0 repositories listed
-
Averaged-DQN: Variance Reduction and Stabilization for Deep Reinforcement Learning7 Nov 2016 0 repositories listed
-
Reinforcement Learning Approach for Parallelization in Filters Aggregation Based Feature Selection Algorithms7 Nov 2016 0 repositories listed
-
Learning to Perform Physics Experiments via Deep Reinforcement Learning6 Nov 2016 0 repositories listed
-
Multi-task learning with deep model based reinforcement learning4 Nov 2016 0 repositories listed
-
Combating Reinforcement Learning's Sisyphean Curse with Intrinsic Fear3 Nov 2016 0 repositories listed
-
Learning Locomotion Skills Using DeepRL: Does the Choice of Action Space Matter?3 Nov 2016 0 repositories listed
-
Quantile Reinforcement Learning3 Nov 2016 0 repositories listed
-
Using a Deep Reinforcement Learning Agent for Traffic Signal Control3 Nov 2016 0 repositories listed
-
Learning Runtime Parameters in Computer Systems with Delayed Experience Injection31 Oct 2016 0 repositories listed
-
Contextual Decision Processes with Low Bellman Rank are PAC-Learnable29 Oct 2016 0 repositories listed
-
Quantum-enhanced machine learning26 Oct 2016 0 repositories listed
-
Reinforcement Learning in Conflicting Environments for Autonomous Vehicles22 Oct 2016 0 repositories listed
-
Utilization of Deep Reinforcement Learning for saccadic-based object visual search20 Oct 2016 0 repositories listed
-
A Reinforcement Learning Approach to the View Planning Problem19 Oct 2016 0 repositories listed
-
Particle Swarm Optimization for Generating Interpretable Fuzzy Reinforcement Learning Policies19 Oct 2016 0 repositories listed
-
Online Contrastive Divergence with Generative Replay: Experience Replay without Storing Data18 Oct 2016 0 repositories listed
-
The End of Optimism? An Asymptotic Analysis of Finite-Armed Linear Bandits14 Oct 2016 0 repositories listed
-
Sim-to-Real Robot Learning from Pixels with Progressive Nets13 Oct 2016 0 repositories listed
-
Introduction to the "Industrial Benchmark"12 Oct 2016 0 repositories listed
-
Navigational Instruction Generation as Inverse Reinforcement Learning with Neural Machine Translation11 Oct 2016 0 repositories listed
-
Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving11 Oct 2016 0 repositories listed
-
Personalizing a Dialogue System with Transfer Reinforcement Learning10 Oct 2016 0 repositories listed
-
Deep Reinforcement Learning From Raw Pixels in Doom7 Oct 2016 0 repositories listed
-
Connecting Generative Adversarial Networks and Actor-Critic Methods6 Oct 2016 0 repositories listed
-
Towards Cognitive Exploration through Deep Reinforcement Learning for Mobile Robots6 Oct 2016 0 repositories listed
-
Reset-Free Guided Policy Search: Efficient Deep Reinforcement Learning with Stochastic Initial States4 Oct 2016 0 repositories listed
-
Collective Robot Reinforcement Learning with Distributed Asynchronous Guided Policy Search3 Oct 2016 0 repositories listed
-
Deep Reinforcement Learning for Robotic Manipulation with Asynchronous Off-Policy Updates3 Oct 2016 0 repositories listed
-
Deep Reinforcement Learning for Tensegrity Robot Locomotion28 Sep 2016 0 repositories listed
-
UbuntuWorld 1.0 LTS - A Platform for Automated Problem Solving & Troubleshooting in the Ubuntu OS27 Sep 2016 0 repositories listed