Browse State-of-the-Art › Reinforcement Learning › Papers, page 98
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 98 of 132: papers 9,701 to 9,800 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
On the Convergence of the Monte Carlo Exploring Starts Algorithm for Reinforcement Learning10 Feb 2020 0 repositories listed
-
Proficiency Constrained Multi-Agent Reinforcement Learning for Environment-Adaptive Multi UAV-UGV Teaming10 Feb 2020 0 repositories listed
-
Provable Self-Play Algorithms for Competitive Reinforcement Learning10 Feb 2020 0 repositories listed
-
Mean-Field Controls with Q-learning for Cooperative MARL: Convergence and Complexity Analysis10 Feb 2020 0 repositories listed
-
Statistically Efficient Off-Policy Policy Gradients10 Feb 2020 0 repositories listed
-
Reward Tweaking: Maximizing the Total Reward While Planning for Short Horizons9 Feb 2020 0 repositories listed
-
A data-driven choice of misfit function for FWI using reinforcement learning8 Feb 2020 0 repositories listed
-
BRPO: Batch Residual Policy Optimization8 Feb 2020 0 repositories listed
-
Capsule Network Performance with Autonomous Navigation8 Feb 2020 0 repositories listed
-
Comprehensive and Efficient Data Labeling via Adaptive Model Scheduling8 Feb 2020 0 repositories listed
-
Conservative Exploration in Reinforcement Learning8 Feb 2020 0 repositories listed
-
Description Based Text Classification with Reinforcement Learning8 Feb 2020 0 repositories listed
-
GLSearch: Maximum Common Subgraph Detection via Learning to Search8 Feb 2020 0 repositories listed
-
Inferential Induction: A Novel Framework for Bayesian Reinforcement Learning8 Feb 2020 0 repositories listed
-
Multi-task Reinforcement Learning with a Planning Quasi-Metric8 Feb 2020 0 repositories listed
-
RL-Duet: Online Music Accompaniment Generation Using Deep Reinforcement Learning8 Feb 2020 0 repositories listed
-
Accelerating Reinforcement Learning for Reaching using Continuous Curriculum Learning7 Feb 2020 0 repositories listed
-
Automated Lane Change Strategy using Proximal Policy Optimization-based Deep Reinforcement Learning7 Feb 2020 0 repositories listed
-
Bayesian Residual Policy Optimization: Scalable Bayesian Reinforcement Learning with Clairvoyant Experts7 Feb 2020 0 repositories listed
-
Causally Correct Partial Models for Reinforcement Learning7 Feb 2020 0 repositories listed
-
Dynamic Energy Dispatch Based on Deep Reinforcement Learning in IoT-Driven Smart Isolated Microgrids7 Feb 2020 0 repositories listed
-
Explicit Mean-Square Error Bounds for Monte-Carlo and Linear Stochastic Approximation7 Feb 2020 0 repositories listed
-
I love your chain mail! Making knights smile in a fantasy game world: Open-domain goal-oriented dialogue agents7 Feb 2020 0 repositories listed
-
Learning Whole-body Motor Skills for Humanoids7 Feb 2020 0 repositories listed
-
Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals7 Feb 2020 0 repositories listed
-
MDLdroid: a ChainSGD-reduce Approach to Mobile Deep Learning for Personal Mobile Sensing7 Feb 2020 0 repositories listed
-
Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)7 Feb 2020 0 repositories listed
-
Representation of Reinforcement Learning Policies in Reproducing Kernel Hilbert Spaces7 Feb 2020 0 repositories listed
-
Ready Policy One: World Building Through Active Learning7 Feb 2020 0 repositories listed
-
Reward-Free Exploration for Reinforcement Learning7 Feb 2020 0 repositories listed
-
Student/Teacher Advising through Reward Augmentation7 Feb 2020 0 repositories listed
-
Reinforcement Learning in Factored MDPs: Oracle-Efficient Algorithms and Tighter Regret Bounds for the Non-Episodic Setting6 Feb 2020 0 repositories listed
-
Social diversity and social preferences in mixed-motive reinforcement learning6 Feb 2020 0 repositories listed
-
Temporal-adaptive Hierarchical Reinforcement Learning6 Feb 2020 0 repositories listed
-
Transfer Heterogeneous Knowledge Among Peer-to-Peer Teammates: A Model Distillation Approach6 Feb 2020 0 repositories listed
-
Unboxing MAC Protocol Design Optimization Using Deep Learning6 Feb 2020 0 repositories listed
-
Deep Learning Tubes for Tube MPC5 Feb 2020 0 repositories listed
-
Deep Radial-Basis Value Functions for Continuous Control5 Feb 2020 0 repositories listed
-
Learning Test-time Augmentation for Content-based Image Retrieval5 Feb 2020 0 repositories listed
-
Mutual Information-based State-Control for Intrinsically Motivated Reinforcement Learning5 Feb 2020 0 repositories listed
-
Bootstrapping a DQN Replay Memory with Synthetic Experiences4 Feb 2020 0 repositories listed
-
Finite Time Analysis of Linear Two-timescale Stochastic Approximation with Markovian Noise4 Feb 2020 0 repositories listed
-
Learning rewards for robotic ultrasound scanning using probabilistic temporal ranking4 Feb 2020 0 repositories listed
-
Learning Task-Driven Control Policies via Information Bottlenecks4 Feb 2020 0 repositories listed
-
Neuro-evolutionary Frameworks for Generalized Learning Agents4 Feb 2020 0 repositories listed
-
Policy Gradient based Quantum Approximate Optimization Algorithm4 Feb 2020 0 repositories listed
-
Evolutionary algorithms for constructing an ensemble of decision trees3 Feb 2020 0 repositories listed
-
Finite-Sample Analysis of Stochastic Approximation Using Smooth Convex Envelopes3 Feb 2020 0 repositories listed
-
Prophet: Proactive Candidate-Selection for Federated Learning by Predicting the Qualities of Training and Reporting Phases3 Feb 2020 0 repositories listed
-
Deep Reinforcement Learning for Autonomous Driving: A Survey2 Feb 2020 0 repositories listed
-
A Deep Reinforcement Learning Approach to Concurrent Bilateral Negotiation31 Jan 2020 0 repositories listed
-
Constrained Deep Reinforcement Learning for Energy Sustainable Multi-UAV based Random Access IoT Networks with NOMA31 Jan 2020 0 repositories listed
-
Locally Private Distributed Reinforcement Learning31 Jan 2020 0 repositories listed
-
Neural MMO v1.3: A Massively Multiagent Game Environment for Training and Evaluating Neural Networks31 Jan 2020 0 repositories listed
-
Predicting Goal-directed Attention Control Using Inverse-Reinforcement Learning31 Jan 2020 0 repositories listed
-
Preventing Imitation Learning with Adversarial Policy Ensembles31 Jan 2020 0 repositories listed
-
Survey of Deep Reinforcement Learning for Motion Planning of Autonomous Vehicles30 Jan 2020 0 repositories listed
-
Asymptotically Efficient Off-Policy Evaluation for Tabular Reinforcement Learning29 Jan 2020 0 repositories listed
-
Multiple Access in Dynamic Cell-Free Networks: Outage Performance and Deep Reinforcement Learning-Based Design29 Jan 2020 0 repositories listed
-
Robust Multimodal Image Registration Using Deep Recurrent Reinforcement Learning29 Jan 2020 0 repositories listed
-
Using Fractal Neural Networks to Play SimCity 1 and Conway's Game of Life at Variable Scales29 Jan 2020 0 repositories listed
-
Variational Autoencoders for Opponent Modeling in Multi-Agent Systems29 Jan 2020 0 repositories listed
-
Data-driven control of micro-climate in buildings: an event-triggered reinforcement learning approach28 Jan 2020 0 repositories listed
-
Distal Explanations for Model-free Explainable Reinforcement Learning28 Jan 2020 0 repositories listed
-
Coagent Networks Revisited28 Jan 2020 0 repositories listed
-
Towards Learning Multi-agent Negotiations via Self-Play28 Jan 2020 0 repositories listed
-
Developing Multi-Task Recommendations with Long-Term Rewards via Policy Distilled Reinforcement Learning27 Jan 2020 0 repositories listed
-
Regret Bounds for Decentralized Learning in Cooperative Multi-Agent Dynamical Systems27 Jan 2020 0 repositories listed
-
Reinforcement Learning-based Application Autoscaling in the Cloud: A Survey27 Jan 2020 0 repositories listed
-
Constrained Upper Confidence Reinforcement Learning26 Jan 2020 0 repositories listed
-
Regime Switching Bandits26 Jan 2020 0 repositories listed
-
Sentiment and Knowledge Based Algorithmic Trading with Deep Reinforcement Learning26 Jan 2020 0 repositories listed
-
Deep Reinforcement Learning based Blind mmWave MIMO Beam Alignment25 Jan 2020 0 repositories listed
-
Following Instructions by Imagining and Reaching Visual Goals25 Jan 2020 0 repositories listed
-
EgoMap: Projective mapping and structured egocentric memory for Deep RL24 Jan 2020 0 repositories listed
-
End-to-End Vision-Based Adaptive Cruise Control (ACC) Using Deep Reinforcement Learning24 Jan 2020 0 repositories listed
-
Stacked Auto Encoder Based Deep Reinforcement Learning for Online Resource Scheduling in Large-Scale MEC Networks24 Jan 2020 0 repositories listed
-
Chameleon: Adaptive Code Optimization for Expedited Deep Neural Network Compilation23 Jan 2020 0 repositories listed
-
Facial Feedback for Reinforcement Learning: A Case Study and Offline Analysis Using the TAMER Framework23 Jan 2020 0 repositories listed
-
Reducing Non-Normative Text Generation from Language Models23 Jan 2020 0 repositories listed
-
Multi-objective Neural Architecture Search via Non-stationary Policy Gradient23 Jan 2020 0 repositories listed
-
Local Policy Optimization for Trajectory-Centric Reinforcement Learning22 Jan 2020 0 repositories listed
-
Machine Learning assisted Handover and Resource Management for Cellular Connected Drones22 Jan 2020 0 repositories listed
-
On Solving Cooperative MARL Problems with a Few Good Experiences22 Jan 2020 0 repositories listed
-
Reinforcement Learning Based Vehicle-cell Association Algorithm for Highly Mobile Millimeter Wave Communication22 Jan 2020 0 repositories listed
-
Cooperative Highway Work Zone Merge Control based on Reinforcement Learning in A Connected and Automated Environment21 Jan 2020 0 repositories listed
-
Improving Interaction Quality Estimation with BiLSTMs and the Impact on Dialogue Policy Learning21 Jan 2020 0 repositories listed
-
Intelligent Bandwidth Allocation for Latency Management in NG-EPON using Reinforcement Learning Methods21 Jan 2020 0 repositories listed
-
Lyceum: An efficient and scalable ecosystem for robot learning21 Jan 2020 0 repositories listed
-
Unsupervisedly Learned Representations: Should the Quest be Over?21 Jan 2020 0 repositories listed
-
A Review on Generative Adversarial Networks: Algorithms, Theory, and Applications20 Jan 2020 0 repositories listed
-
Memristor Hardware-Friendly Reinforcement Learning20 Jan 2020 0 repositories listed
-
Nested-Wasserstein Self-Imitation Learning for Sequence Generation20 Jan 2020 0 repositories listed
-
Reinforcement Learning with Probabilistically Complete Exploration20 Jan 2020 0 repositories listed
-
A Survey of Reinforcement Learning Techniques: Strategies, Recent Development, and Future Directions19 Jan 2020 0 repositories listed
-
FRESH: Interactive Reward Shaping in High-Dimensional State Spaces using Human Feedback19 Jan 2020 0 repositories listed
-
Learning Options from Demonstration using Skill Segmentation19 Jan 2020 0 repositories listed
-
cube2net: Efficient Query-Specific Network Construction with Data Cube Organization18 Jan 2020 0 repositories listed
-
Effects of sparse rewards of different magnitudes in the speed of learning of model-based actor critic methods18 Jan 2020 0 repositories listed
-
BNAS:An Efficient Neural Architecture Search Approach Using Broad Scalable Architecture18 Jan 2020 0 repositories listed