Browse State-of-the-Art › Reinforcement Learning › Papers, page 108
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 108 of 132: papers 10,701 to 10,800 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
What Should I Ask? Using Conversationally Informative Rewards for Goal-oriented Visual Dialog.1 Jul 2019 0 repositories listed
-
Collaboration of AI Agents via Cooperative Multi-Agent Deep Reinforcement Learning30 Jun 2019 0 repositories listed
-
Machine Learning for Intelligent Authentication in 5G-and-Beyond Wireless Networks30 Jun 2019 0 repositories listed
-
Multi-Armed Bandits with Fairness Constraints for Distributing Resources to Human Teammates30 Jun 2019 0 repositories listed
-
On Training Flexible Robots using Deep Reinforcement Learning29 Jun 2019 0 repositories listed
-
Learning to Cope with Adversarial Attacks28 Jun 2019 0 repositories listed
-
Adaptive Honeypot Engagement through Reinforcement Learning of Semi-Markov Decision Processes27 Jun 2019 0 repositories listed
-
Demonstration-Guided Deep Reinforcement Learning of Control Policies for Dexterous Human-Robot Interaction27 Jun 2019 0 repositories listed
-
ExTra: Transfer-guided Exploration27 Jun 2019 0 repositories listed
-
From self-tuning regulators to reinforcement learning and back again27 Jun 2019 0 repositories listed
-
Learning Policies through Quantile Regression27 Jun 2019 0 repositories listed
-
Toward Simulating Environments in Reinforcement Learning Based Recommendations27 Jun 2019 0 repositories listed
-
A Tractable Algorithm For Finite-Horizon Continuous Reinforcement Learning26 Jun 2019 0 repositories listed
-
Approximate Dynamic Programming For Linear Systems with State and Input Constraints26 Jun 2019 0 repositories listed
-
Efficient Navigation of Colloidal Robots in an Unknown Environment via Deep Reinforcement Learning26 Jun 2019 0 repositories listed
-
Compositional Transfer in Hierarchical Reinforcement Learning26 Jun 2019 0 repositories listed
-
Rethinking Formal Models of Partially Observable Multiagent Decision Making26 Jun 2019 0 repositories listed
-
Expected Sarsa(λ) with Control Variate for Variance Reduction25 Jun 2019 0 repositories listed
-
Learning Causal State Representations of Partially Observable Environments25 Jun 2019 0 repositories listed
-
Neural Proximal/Trust Region Policy Optimization Attains Globally Optimal Policy25 Jun 2019 0 repositories listed
-
On Multi-Agent Learning in Team Sports Games25 Jun 2019 0 repositories listed
-
Optimistic Proximal Policy Optimization25 Jun 2019 0 repositories listed
-
Policy Optimization with Stochastic Mirror Descent25 Jun 2019 0 repositories listed
-
Reinforcement Learning with Competitive Ensembles of Information-Constrained Primitives25 Jun 2019 0 repositories listed
-
Uncertainty-aware Model-based Policy Optimization25 Jun 2019 0 repositories listed
-
A Theoretical Connection Between Statistical Physics and Reinforcement Learning24 Jun 2019 0 repositories listed
-
Deceptive Reinforcement Learning Under Adversarial Manipulations on Cost Signals24 Jun 2019 0 repositories listed
-
Deep Conservative Policy Iteration24 Jun 2019 0 repositories listed
-
Event-Driven Models24 Jun 2019 0 repositories listed
-
Inverse reinforcement learning conditioned on brain scan24 Jun 2019 0 repositories listed
-
Optimal Use of Experience in First Person Shooter Environments24 Jun 2019 0 repositories listed
-
Neural networks with motivation23 Jun 2019 0 repositories listed
-
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference23 Jun 2019 0 repositories listed
-
Reinforcement Learning-Based Trajectory Design for the Aerial Base Stations23 Jun 2019 0 repositories listed
-
RLTM: An Efficient Neural IR Framework for Long Documents22 Jun 2019 0 repositories listed
-
A Study of State Aliasing in Structured Prediction with RNNs21 Jun 2019 0 repositories listed
-
Continual Reinforcement Learning with Diversity Exploration and Adversarial Self-Correction21 Jun 2019 0 repositories listed
-
Disentangled Skill Embeddings for Reinforcement Learning21 Jun 2019 0 repositories listed
-
Leveraging Reinforcement Learning Techniques for Effective Policy Adoption and Validation21 Jun 2019 0 repositories listed
-
Revised Progressive-Hedging-Algorithm Based Two-layer Solution Scheme for Bayesian Reinforcement Learning21 Jun 2019 0 repositories listed
-
Shaping Belief States with Generative Environment Models for RL21 Jun 2019 0 repositories listed
-
Cache-Aided NOMA Mobile Edge Computing: A Reinforcement Learning Approach20 Jun 2019 0 repositories listed
-
Cooperative Lane Changing via Deep Reinforcement Learning20 Jun 2019 0 repositories listed
-
Finding Needles in a Moving Haystack: Prioritizing Alerts with Adversarial Reinforcement Learning20 Jun 2019 0 repositories listed
-
The Finite-Horizon Two-Armed Bandit Problem with Binary Responses: A Multidisciplinary Survey of the History, State of the Art, and Myths20 Jun 2019 0 repositories listed
-
Variable Impedance Control in End-Effector Space: An Action Space for Reinforcement Learning in Contact-Rich Tasks20 Jun 2019 0 repositories listed
-
When Multiple Agents Learn to Schedule: A Distributed Radio Resource Management Framework20 Jun 2019 0 repositories listed
-
Adapting Behaviour via Intrinsic Reward: A Survey and Empirical Study19 Jun 2019 0 repositories listed
-
Adaptive Temporal-Difference Learning for Policy Evaluation with Per-State Uncertainty Estimates19 Jun 2019 0 repositories listed
-
Experience Replay Optimization19 Jun 2019 0 repositories listed
-
Global Convergence of Policy Gradient Methods to (Almost) Locally Optimal Policies19 Jun 2019 0 repositories listed
-
Multi-user Resource Control with Deep Reinforcement Learning in IoT Edge Computing19 Jun 2019 0 repositories listed
-
Reward Prediction Error as an Exploration Objective in Deep RL19 Jun 2019 0 repositories listed
-
Wasserstein Adversarial Imitation Learning19 Jun 2019 0 repositories listed
-
Directed Exploration for Reinforcement Learning18 Jun 2019 0 repositories listed
-
DISCO: Influence Maximization Meets Network Embedding and Deep Learning18 Jun 2019 0 repositories listed
-
Evolutionary Reinforcement Learning for Sample-Efficient Multiagent Coordination18 Jun 2019 0 repositories listed
-
Gap-Increasing Policy Evaluation for Efficient and Noise-Tolerant Reinforcement Learning18 Jun 2019 0 repositories listed
-
Hill Climbing on Value Estimates for Search-control in Dyna18 Jun 2019 0 repositories listed
-
RIDM: Reinforced Inverse Dynamics Modeling for Learning from a Single Observed Demonstration18 Jun 2019 0 repositories listed
-
Robust Reinforcement Learning for Continuous Control with Model Misspecification18 Jun 2019 0 repositories listed
-
Sample-efficient Adversarial Imitation Learning from Observation18 Jun 2019 0 repositories listed
-
Towards White-box Benchmarks for Algorithm Control18 Jun 2019 0 repositories listed
-
A gray-box approach for curriculum learning17 Jun 2019 0 repositories listed
-
Bayesian Optimization with Binary Auxiliary Information17 Jun 2019 0 repositories listed
-
Iterative Model-Based Reinforcement Learning Using Simulations in the Differentiable Neural Computer17 Jun 2019 0 repositories listed
-
LPaintB: Learning to Paint from Self-Supervision17 Jun 2019 0 repositories listed
-
A Joint Planning and Learning Framework for Human-Aided Decision-Making17 Jun 2019 0 repositories listed
-
Universal Successor Features Based Deep Reinforcement Learning for Navigation17 Jun 2019 0 repositories listed
-
Reinforcement Learning Driven Heuristic Optimization16 Jun 2019 0 repositories listed
-
Injecting Prior Knowledge for Transfer Learning into Reinforcement Learning Algorithms using Logic Tensor Networks15 Jun 2019 0 repositories listed
-
Reinforcement Learning with Non-uniform State Representations for Adaptive Search15 Jun 2019 0 repositories listed
-
Direct Policy Gradients: Direct Optimization of Policies in Discrete Action Spaces14 Jun 2019 0 repositories listed
-
Epistemic Risk-Sensitive Reinforcement Learning14 Jun 2019 0 repositories listed
-
Provably Efficient Q-learning with Function Approximation via Distribution Shift Error Checking Oracle14 Jun 2019 0 repositories listed
-
Self-Tuning Sectorization: Deep Reinforcement Learning Meets Broadcast Beam Optimization14 Jun 2019 0 repositories listed
-
Deep Reinforcement Learning for Cyber Security13 Jun 2019 0 repositories listed
-
Early Detection of Long Term Evaluation Criteria in Online Controlled Experiments13 Jun 2019 0 repositories listed
-
Conditioning of Reinforcement Learning Agents and its Policy Regularization Application13 Jun 2019 0 repositories listed
-
Modeling and Interpreting Real-world Human Risk Decision Making with Inverse Reinforcement Learning13 Jun 2019 0 repositories listed
-
Sub-policy Adaptation for Hierarchical Reinforcement Learning13 Jun 2019 0 repositories listed
-
Deep Reinforcement Learning for Unmanned Aerial Vehicle-Assisted Vehicular Networks12 Jun 2019 0 repositories listed
-
Regret Minimization for Reinforcement Learning by Evaluating the Optimal Bias Function12 Jun 2019 0 repositories listed
-
Adaptive Optimal Control for Reference Tracking Independent of Exo-System Dynamics12 Jun 2019 0 repositories listed
-
Sub-Goal Trees -- a Framework for Goal-Directed Trajectory Prediction and Optimization12 Jun 2019 0 repositories listed
-
A Hybrid Approach Between Adversarial Generative Networks and Actor-Critic Policy Gradient for Low Rate High-Resolution Image Compression11 Jun 2019 0 repositories listed
-
Continual Reinforcement Learning deployed in Real-life using Policy Distillation and Sim2Real Transfer11 Jun 2019 0 repositories listed
-
Dealing with Non-Stationarity in Multi-Agent Deep Reinforcement Learning11 Jun 2019 0 repositories listed
-
Deep learning control of artificial avatars in group coordination tasks11 Jun 2019 0 repositories listed
-
Reinforcement Learning for Integer Programming: Learning to Cut11 Jun 2019 0 repositories listed
-
Reinforcement Learning of Minimalist Numeral Grammars11 Jun 2019 0 repositories listed
-
Towards Inverse Reinforcement Learning for Limit Order Book Dynamics11 Jun 2019 0 repositories listed
-
A Survey of Reinforcement Learning Informed by Natural Language10 Jun 2019 0 repositories listed
-
Attacking Graph Convolutional Networks via Rewiring10 Jun 2019 0 repositories listed
-
Deep Reinforcement Learning with Discrete Normalized Advantage Functions for Resource Management in Network Slicing10 Jun 2019 0 repositories listed
-
Model-Based Reinforcement Learning with a Generative Model is Minimax Optimal10 Jun 2019 0 repositories listed
-
Neural Heterogeneous Scheduler9 Jun 2019 0 repositories listed
-
Transfer Learning by Modeling a Distribution over Policies9 Jun 2019 0 repositories listed
-
Towards Optimal Off-Policy Evaluation for Reinforcement Learning with Marginalized Importance Sampling8 Jun 2019 0 repositories listed
-
Figure Captioning with Reasoning and Sequence-Level Training7 Jun 2019 0 repositories listed