Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 130
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 130 of 152: papers 12,901 to 13,000 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
City Metro Network Expansion with Reinforcement Learning25 Sep 2019 0 repositories listed
-
Collaborative Inter-agent Knowledge Distillation for Reinforcement Learning25 Sep 2019 0 repositories listed
-
Consistent Meta-Reinforcement Learning via Model Identification and Experience Relabeling25 Sep 2019 0 repositories listed
-
Contextual Inverse Reinforcement Learning25 Sep 2019 0 repositories listed
-
Counterfactual Regularization for Model-Based Reinforcement Learning25 Sep 2019 0 repositories listed
-
CrossNorm: On Normalization for Off-Policy Reinforcement Learning25 Sep 2019 0 repositories listed
-
Deep RL for Blood Glucose Control: Lessons, Challenges, and Opportunities25 Sep 2019 0 repositories listed
-
DeepAGREL: Biologically plausible deep learning via direct reinforcement25 Sep 2019 0 repositories listed
-
Do recent advancements in model-based deep reinforcement learning really improve data efficiency?25 Sep 2019 0 repositories listed
-
Efficient meta reinforcement learning via meta goal generation25 Sep 2019 0 repositories listed
-
Event Discovery for History Representation in Reinforcement Learning25 Sep 2019 0 repositories listed
-
Evo-NAS: Evolutionary-Neural Hybrid Agent for Architecture Search25 Sep 2019 0 repositories listed
-
Generalizing Reinforcement Learning to Unseen Actions25 Sep 2019 0 repositories listed
-
HIPPOCAMPAL NEURONAL REPRESENTATIONS IN CONTINUAL LEARNING25 Sep 2019 0 repositories listed
-
Hope For The Best But Prepare For The Worst: Cautious Adaptation In RL Agents25 Sep 2019 0 repositories listed
-
How many weights are enough : can tensor factorization learn efficient policies ?25 Sep 2019 0 repositories listed
-
Improving Exploration of Deep Reinforcement Learning using Planning for Policy Search25 Sep 2019 0 repositories listed
-
Improving SAT Solver Heuristics with Graph Networks and Reinforcement Learning25 Sep 2019 0 repositories listed
-
Learning Algorithmic Solutions to Symbolic Planning Tasks with a Neural Computer25 Sep 2019 0 repositories listed
-
Learning by shaking: Computing policy gradients by physical forward-propagation25 Sep 2019 0 repositories listed
-
Learning Functionally Decomposed Hierarchies for Continuous Navigation Tasks25 Sep 2019 0 repositories listed
-
Learning Good Policies By Learning Good Perceptual Models25 Sep 2019 0 repositories listed
-
Learning Key Steps to Attack Deep Reinforcement Learning Agents25 Sep 2019 0 repositories listed
-
Learning Temporal Abstraction with Information-theoretic Constraints for Hierarchical Reinforcement Learning25 Sep 2019 0 repositories listed
-
Learning to Reach Goals Without Reinforcement Learning25 Sep 2019 0 repositories listed
-
Learning to Reason: Distilling Hierarchy via Self-Supervision and Reinforcement Learning25 Sep 2019 0 repositories listed
-
Learning with Social Influence through Interior Policy Differentiation25 Sep 2019 0 repositories listed
-
Learning World Graph Decompositions To Accelerate Reinforcement Learning25 Sep 2019 0 repositories listed
-
Long-term planning, short-term adjustments25 Sep 2019 0 repositories listed
-
Meta Learning via Learned Loss25 Sep 2019 0 repositories listed
-
Mint: Matrix-Interleaving for Multi-Task Learning25 Sep 2019 0 repositories listed
-
Model Ensemble-Based Intrinsic Reward for Sparse Reward Reinforcement Learning25 Sep 2019 0 repositories listed
-
Model-free Learning Control of Nonlinear Stochastic Systems with Stability Guarantee25 Sep 2019 0 repositories listed
-
Model Imitation for Model-Based Reinforcement Learning25 Sep 2019 0 repositories listed
-
Modeling Fake News in Social Networks with Deep Multi-Agent Reinforcement Learning25 Sep 2019 0 repositories listed
-
MoET: Interpretable and Verifiable Reinforcement Learning via Mixture of Expert Trees25 Sep 2019 0 repositories listed
-
Multi-Agent Hierarchical Reinforcement Learning for Humanoid Navigation25 Sep 2019 0 repositories listed
-
Multi-step Greedy Policies in Model-Free Deep Reinforcement Learning25 Sep 2019 0 repositories listed
-
Multiagent Reinforcement Learning in Games with an Iterated Dominance Solution25 Sep 2019 0 repositories listed
-
Partial Simulation for Imitation Learning25 Sep 2019 0 repositories listed
-
Policy Optimization by Local Improvement through Search25 Sep 2019 0 repositories listed
-
Policy Tree Network25 Sep 2019 0 repositories listed
-
Multi-task Batch Reinforcement Learning with Metric Learning25 Sep 2019 0 repositories listed
-
Pre-training as Batch Meta Reinforcement Learning with tiMe25 Sep 2019 0 repositories listed
-
Probabilistic View of Multi-agent Reinforcement Learning: A Unified Approach25 Sep 2019 0 repositories listed
-
QXplore: Q-Learning Exploration by Maximizing Temporal Difference Error25 Sep 2019 0 repositories listed
-
REFINING MONTE CARLO TREE SEARCH AGENTS BY MONTE CARLO TREE SEARCH25 Sep 2019 0 repositories listed
-
Reinforcement learning for suppression of collective activity in oscillatory ensembles25 Sep 2019 0 repositories listed
-
Reinforcement Learning with Chromatic Networks25 Sep 2019 0 repositories listed
-
Risk Averse Value Expansion for Sample Efficient and Robust Policy Learning25 Sep 2019 0 repositories listed
-
Robust Domain Randomization for Reinforcement Learning25 Sep 2019 0 repositories listed
-
S2VG: Soft Stochastic Value Gradient method25 Sep 2019 0 repositories listed
-
Sequence-level Intrinsic Exploration Model for Partially Observable Domains25 Sep 2019 0 repositories listed
-
Solving single-objective tasks by preference multi-objective reinforcement learning25 Sep 2019 0 repositories listed
-
Sparse Skill Coding: Learning Behavioral Hierarchies with Sparse Codes25 Sep 2019 0 repositories listed
-
Stabilizing Off-Policy Reinforcement Learning with Conservative Policy Gradients25 Sep 2019 0 repositories listed
-
Striving for Simplicity in Off-Policy Deep Reinforcement Learning25 Sep 2019 0 repositories listed
-
Subjective Reinforcement Learning for Open Complex Environments25 Sep 2019 0 repositories listed
-
Temporal Difference Weighted Ensemble For Reinforcement Learning25 Sep 2019 0 repositories listed
-
Towards Simplicity in Deep Reinforcement Learning: Streamlined Off-Policy Learning25 Sep 2019 0 repositories listed
-
Training a Constrained Natural Media Painting Agent using Reinforcement Learning25 Sep 2019 0 repositories listed
-
Trajectory representation learning for Multi-Task NMRDPs planning25 Sep 2019 0 repositories listed
-
Variational Constrained Reinforcement Learning with Application to Planning at Roundabout25 Sep 2019 0 repositories listed
-
Zero-Shot Policy Transfer with Disentangled Attention25 Sep 2019 0 repositories listed
-
Brain-Inspired Hardware for Artificial Intelligence: Accelerated Learning in a Physical-Model Spiking Neural Network24 Sep 2019 0 repositories listed
-
Controlling an Autonomous Vehicle with Deep Reinforcement Learning24 Sep 2019 0 repositories listed
-
Efficient Inference and Exploration for Reinforcement Learning24 Sep 2019 0 repositories listed
-
Power Allocation in Cache-Aided NOMA Systems: Optimization and Deep Reinforcement Learning Approaches24 Sep 2019 0 repositories listed
-
Where to Look Next: Unsupervised Active Visual Exploration on 360° Input23 Sep 2019 0 repositories listed
-
Robot Navigation in Crowds by Graph Convolutional Networks with Attention Learned from Human Gaze23 Sep 2019 0 repositories listed
-
PAC Reinforcement Learning without Real-World Feedback23 Sep 2019 0 repositories listed
-
Constrained Attractor Selection Using Deep Reinforcement Learning23 Sep 2019 0 repositories listed
-
Integrating independent and centralized multi-agent reinforcement learning for traffic signal network optimization23 Sep 2019 0 repositories listed
-
Why Does Hierarchy (Sometimes) Work So Well in Reinforcement Learning?23 Sep 2019 0 repositories listed
-
Leveraging Human Guidance for Deep Reinforcement Learning Tasks21 Sep 2019 0 repositories listed
-
A Layered Architecture for Active Perception: Image Classification using Deep Reinforcement Learning20 Sep 2019 0 repositories listed
-
How Much Do Unstated Problem Constraints Limit Deep Robotic Reinforcement Learning?20 Sep 2019 0 repositories listed
-
On the Convergence of Approximate and Regularized Policy Iteration Schemes20 Sep 2019 0 repositories listed
-
Redirection Controller Using Reinforcement Learning20 Sep 2019 0 repositories listed
-
MACS: Deep Reinforcement Learning based SDN Controller Synchronization Policy Design19 Sep 2019 0 repositories listed
-
Robot Sound Interpretation: Combining Sight and Sound in Learning-Based Control19 Sep 2019 0 repositories listed
-
Instance-dependent ℓ_∞-bounds for policy evaluation in tabular reinforcement learning19 Sep 2019 0 repositories listed
-
A Hierarchical Two-tier Approach to Hyper-parameter Optimization in Reinforcement Learning18 Sep 2019 0 repositories listed
-
A Human-Centered Data-Driven Planner-Actor-Critic Architecture via Logic Programming18 Sep 2019 0 repositories listed
-
Automated Lane Change Decision Making using Deep Reinforcement Learning in Dynamic and Uncertain Highway Environment18 Sep 2019 0 repositories listed
-
DeepGait: Planning and Control of Quadrupedal Gaits using Deep Reinforcement Learning18 Sep 2019 0 repositories listed
-
Dependency-Aware Computation Offloading in Mobile Edge Computing: A Reinforcement Learning Approach18 Sep 2019 0 repositories listed
-
Robust Opponent Modeling via Adversarial Ensemble Reinforcement Learning in Asymmetric Imperfect-Information Games18 Sep 2019 0 repositories listed
-
Segregation Dynamics with Reinforcement Learning and Agent Based Modeling18 Sep 2019 0 repositories listed
-
Visual Tracking by means of Deep Reinforcement Learning and an Expert Demonstrator18 Sep 2019 0 repositories listed
-
A Review of Tracking, Prediction and Decision Making Methods for Autonomous Driving17 Sep 2019 0 repositories listed
-
Adversarial Feature Training for Generalizable Robotic Visuomotor Control17 Sep 2019 0 repositories listed
-
Attraction-Repulsion Actor-Critic for Continuous Control Reinforcement Learning17 Sep 2019 0 repositories listed
-
Controllable Length Control Neural Encoder-Decoder via Reinforcement Learning17 Sep 2019 0 repositories listed
-
Generating Black-Box Adversarial Examples for Text Classifiers Using a Deep Reinforced Model17 Sep 2019 0 repositories listed
-
Stock market microstructure inference via multi-agent reinforcement learning17 Sep 2019 0 repositories listed
-
Selective Network Discovery via Deep Reinforcement Learning on Embedded Spaces16 Sep 2019 0 repositories listed
-
Data Centers Job Scheduling with Deep Reinforcement Learning16 Sep 2019 0 repositories listed
-
Leveraging human Domain Knowledge to model an empirical Reward function for a Reinforcement Learning problem16 Sep 2019 0 repositories listed
-
Meta Reinforcement Learning for Sim-to-real Domain Adaptation16 Sep 2019 0 repositories listed