Browse State-of-the-Art › reinforcement-learning › Papers, page 112
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 112 of 135: papers 11,101 to 11,200 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Multi-Issue Bargaining With Deep Reinforcement Learning18 Feb 2020 0 repositories listed
-
Reinforcement learning for the privacy preservation and manipulation of eye tracking data17 Feb 2020 0 repositories listed
-
Reward Design for Driver Repositioning Using Multi-Agent Reinforcement Learning17 Feb 2020 0 repositories listed
-
Investigating Simple Object Representations in Model-Free Deep Reinforcement Learning16 Feb 2020 0 repositories listed
-
Non-asymptotic Convergence of Adam-type Reinforcement Learning Algorithms under Markovian Sampling15 Feb 2020 0 repositories listed
-
The Archimedean trap: Why traditional reinforcement learning will probably not yield AGI15 Feb 2020 0 repositories listed
-
Resource Management in Wireless Networks via Multi-Agent Deep Reinforcement Learning14 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning-Based Beam Tracking for Low-Latency Services in Vehicular Networks13 Feb 2020 0 repositories listed
-
Fast Reinforcement Learning for Anti-jamming Communications13 Feb 2020 0 repositories listed
-
Improving Generalization of Reinforcement Learning with Minimax Distributional Soft Actor-Critic13 Feb 2020 0 repositories listed
-
MODRL/D-AM: Multiobjective Deep Reinforcement Learning Algorithm Using Decomposition and Attention Model for Multiobjective Optimization13 Feb 2020 0 repositories listed
-
Multi-Vehicle Routing Problems with Soft Time Windows: A Multi-Agent Reinforcement Learning Approach13 Feb 2020 0 repositories listed
-
Data Efficient Training for Reinforcement Learning with Adaptive Behavior Policy Sharing12 Feb 2020 0 repositories listed
-
Confounding-Robust Policy Evaluation in Infinite-Horizon Reinforcement Learning11 Feb 2020 0 repositories listed
-
HMRL: Hyper-Meta Learning for Sparse Reward Reinforcement Learning Problem11 Feb 2020 0 repositories listed
-
Learning Structured Communication for Multi-agent Reinforcement Learning11 Feb 2020 0 repositories listed
-
Learning to Switch Among Agents in a Team via 2-Layer Markov Decision Processes11 Feb 2020 0 repositories listed
-
Machine Learning Approaches For Motor Learning: A Short Review11 Feb 2020 0 repositories listed
-
Towards Intelligent Pick and Place Assembly of Individualized Products Using Reinforcement Learning11 Feb 2020 0 repositories listed
-
Interpretable Off-Policy Evaluation in Reinforcement Learning by Highlighting Influential Transitions10 Feb 2020 0 repositories listed
-
On the Convergence of the Monte Carlo Exploring Starts Algorithm for Reinforcement Learning10 Feb 2020 0 repositories listed
-
Proficiency Constrained Multi-Agent Reinforcement Learning for Environment-Adaptive Multi UAV-UGV Teaming10 Feb 2020 0 repositories listed
-
Provable Self-Play Algorithms for Competitive Reinforcement Learning10 Feb 2020 0 repositories listed
-
Reward Tweaking: Maximizing the Total Reward While Planning for Short Horizons9 Feb 2020 0 repositories listed
-
A data-driven choice of misfit function for FWI using reinforcement learning8 Feb 2020 0 repositories listed
-
BRPO: Batch Residual Policy Optimization8 Feb 2020 0 repositories listed
-
Conservative Exploration in Reinforcement Learning8 Feb 2020 0 repositories listed
-
Description Based Text Classification with Reinforcement Learning8 Feb 2020 0 repositories listed
-
Inferential Induction: A Novel Framework for Bayesian Reinforcement Learning8 Feb 2020 0 repositories listed
-
Multi-task Reinforcement Learning with a Planning Quasi-Metric8 Feb 2020 0 repositories listed
-
RL-Duet: Online Music Accompaniment Generation Using Deep Reinforcement Learning8 Feb 2020 0 repositories listed
-
Accelerating Reinforcement Learning for Reaching using Continuous Curriculum Learning7 Feb 2020 0 repositories listed
-
Automated Lane Change Strategy using Proximal Policy Optimization-based Deep Reinforcement Learning7 Feb 2020 0 repositories listed
-
Bayesian Residual Policy Optimization: Scalable Bayesian Reinforcement Learning with Clairvoyant Experts7 Feb 2020 0 repositories listed
-
Causally Correct Partial Models for Reinforcement Learning7 Feb 2020 0 repositories listed
-
Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals7 Feb 2020 0 repositories listed
-
Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces7 Feb 2020 0 repositories listed
-
Reward-Free Exploration for Reinforcement Learning7 Feb 2020 0 repositories listed
-
Student/Teacher Advising through Reward Augmentation7 Feb 2020 0 repositories listed
-
Reinforcement Learning in Factored MDPs: Oracle-Efficient Algorithms and Tighter Regret Bounds for the Non-Episodic Setting6 Feb 2020 0 repositories listed
-
Social diversity and social preferences in mixed-motive reinforcement learning6 Feb 2020 0 repositories listed
-
Temporal-adaptive Hierarchical Reinforcement Learning6 Feb 2020 0 repositories listed
-
Mutual Information-based State-Control for Intrinsically Motivated Reinforcement Learning5 Feb 2020 0 repositories listed
-
Learning Task-Driven Control Policies via Information Bottlenecks4 Feb 2020 0 repositories listed
-
Evolutionary algorithms for constructing an ensemble of decision trees3 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning for Autonomous Driving: A Survey2 Feb 2020 0 repositories listed
-
A Deep Reinforcement Learning Approach to Concurrent Bilateral Negotiation31 Jan 2020 0 repositories listed
-
Locally Private Distributed Reinforcement Learning31 Jan 2020 0 repositories listed
-
Predicting Goal-directed Attention Control Using Inverse-Reinforcement Learning31 Jan 2020 0 repositories listed
-
Preventing Imitation Learning with Adversarial Policy Ensembles31 Jan 2020 0 repositories listed
-
Survey of Deep Reinforcement Learning for Motion Planning of Autonomous Vehicles30 Jan 2020 0 repositories listed
-
Asymptotically Efficient Off-Policy Evaluation for Tabular Reinforcement Learning29 Jan 2020 0 repositories listed
-
Robust Multimodal Image Registration Using Deep Recurrent Reinforcement Learning29 Jan 2020 0 repositories listed
-
Distal Explanations for Model-free Explainable Reinforcement Learning28 Jan 2020 0 repositories listed
-
Coagent Networks Revisited28 Jan 2020 0 repositories listed
-
Developing Multi-Task Recommendations with Long-Term Rewards via Policy Distilled Reinforcement Learning27 Jan 2020 0 repositories listed
-
Reinforcement Learning-based Application Autoscaling in the Cloud: A Survey27 Jan 2020 0 repositories listed
-
Constrained Upper Confidence Reinforcement Learning26 Jan 2020 0 repositories listed
-
Sentiment and Knowledge Based Algorithmic Trading with Deep Reinforcement Learning26 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning based Blind mmWave MIMO Beam Alignment25 Jan 2020 0 repositories listed
-
EgoMap: Projective mapping and structured egocentric memory for Deep RL24 Jan 2020 0 repositories listed
-
End-to-End Vision-Based Adaptive Cruise Control (ACC) Using Deep Reinforcement Learning24 Jan 2020 0 repositories listed
-
Pricing commodity swing options24 Jan 2020 0 repositories listed
-
Facial Feedback for Reinforcement Learning: A Case Study and Offline Analysis Using the TAMER Framework23 Jan 2020 0 repositories listed
-
Reducing Non-Normative Text Generation from Language Models23 Jan 2020 0 repositories listed
-
Local Policy Optimization for Trajectory-Centric Reinforcement Learning22 Jan 2020 0 repositories listed
-
On Solving Cooperative MARL Problems with a Few Good Experiences22 Jan 2020 0 repositories listed
-
Intelligent Bandwidth Allocation for Latency Management in NG-EPON using Reinforcement Learning Methods21 Jan 2020 0 repositories listed
-
Lyceum: An efficient and scalable ecosystem for robot learning21 Jan 2020 0 repositories listed
-
Unsupervisedly Learned Representations: Should the Quest be Over?21 Jan 2020 0 repositories listed
-
Memristor Hardware-Friendly Reinforcement Learning20 Jan 2020 0 repositories listed
-
Nested-Wasserstein Self-Imitation Learning for Sequence Generation20 Jan 2020 0 repositories listed
-
Reinforcement Learning with Probabilistically Complete Exploration20 Jan 2020 0 repositories listed
-
A Survey of Reinforcement Learning Techniques: Strategies, Recent Development, and Future Directions19 Jan 2020 0 repositories listed
-
FRESH: Interactive Reward Shaping in High-Dimensional State Spaces using Human Feedback19 Jan 2020 0 repositories listed
-
Learning Options from Demonstration using Skill Segmentation19 Jan 2020 0 repositories listed
-
cube2net: Efficient Query-Specific Network Construction with Data Cube Organization18 Jan 2020 0 repositories listed
-
BNAS:An Efficient Neural Architecture Search Approach Using Broad Scalable Architecture18 Jan 2020 0 repositories listed
-
Algorithms in Multi-Agent Systems: A Holistic Perspective from Reinforcement Learning and Game Theory17 Jan 2020 0 repositories listed
-
MIME: Mutual Information Minimisation Exploration16 Jan 2020 0 repositories listed
-
Reward Shaping for Reinforcement Learning with Omega-Regular Objectives16 Jan 2020 0 repositories listed
-
Model-based Multi-Agent Reinforcement Learning with Cooperative Prioritized Sweeping15 Jan 2020 0 repositories listed
-
Robotic Grasp Manipulation Using Evolutionary Computing and Deep Reinforcement Learning15 Jan 2020 0 repositories listed
-
SEERL: Sample Efficient Ensemble Reinforcement Learning15 Jan 2020 0 repositories listed
-
Exploiting Language Instructions for Interpretable and Compositional Reinforcement Learning13 Jan 2020 0 repositories listed
-
Multi-Robot Formation Control Using Reinforcement Learning13 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning for Complex Manipulation Tasks with Sparse Feedback12 Jan 2020 0 repositories listed
-
Weakly Supervised Video Summarization by Hierarchical Reinforcement Learning12 Jan 2020 0 repositories listed
-
A storage expansion planning framework using reinforcement learning and simulation-based optimization10 Jan 2020 0 repositories listed
-
Deep Interactive Reinforcement Learning for Path Following of Autonomous Underwater Vehicle10 Jan 2020 0 repositories listed
-
Reinforcement Learning Tracking Control for Robotic Manipulator With Kernel-Based Dynamic Model9 Jan 2020 0 repositories listed
-
EEG-based Drowsiness Estimation for Driving Safety using Deep Q-Learning8 Jan 2020 0 repositories listed
-
Multi-Agent Deep Reinforcement Learning for Cooperative Connected Vehicles8 Jan 2020 0 repositories listed
-
On Thompson Sampling for Smoother-than-Lipschitz Bandits8 Jan 2020 0 repositories listed
-
Perception and Navigation in Autonomous Systems in the Era of Learning: A Survey8 Jan 2020 0 repositories listed
-
Sample-based Distributional Policy Gradient8 Jan 2020 0 repositories listed
-
Decentralized Automotive Radar Spectrum Allocation to Avoid Mutual Interference Using Reinforcement Learning7 Jan 2020 0 repositories listed
-
Experimental Analysis of Reinforcement Learning Techniques for Spectrum Sharing Radar6 Jan 2020 0 repositories listed
-
High-speed Autonomous Drifting with Deep Reinforcement Learning6 Jan 2020 0 repositories listed
-
Generalizing Emergent Communication6 Jan 2020 0 repositories listed