Browse State-of-the-Art › reinforcement-learning › Papers, page 85
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 85 of 135: papers 8,401 to 8,500 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Differentially Private Reinforcement Learning with Linear Function Approximation18 Jan 2022 0 repositories listed
-
K-nearest Multi-agent Deep Reinforcement Learning for Collaborative Tasks with a Variable Number of Agents18 Jan 2022 0 repositories listed
-
Programmatic Policy Extraction by Iterative Local Search18 Jan 2022 0 repositories listed
-
Toward Self-learning End-to-End Task-Oriented Dialog Systems18 Jan 2022 0 repositories listed
-
11 Summaries of Papers on Explainable Reinforcement Learning With Some Commentary17 Jan 2022 0 repositories listed
-
An Improved Reinforcement Learning Algorithm for Learning to Branch17 Jan 2022 0 repositories listed
-
An Understanding of Learning from Demonstrations for Neural Text Generation17 Jan 2022 0 repositories listed
-
Chaining Value Functions for Off-Policy Learning17 Jan 2022 0 repositories listed
-
Exploration by Random Network Distillation17 Jan 2022 0 repositories listed
-
Implementations that Matter in Cooperative Multi-Agent Reinforcement Learning17 Jan 2022 0 repositories listed
-
Railway Operation Rescheduling System via Dynamic Simulation and Reinforcement Learning17 Jan 2022 0 repositories listed
-
Spatiotemporal Costmap Inference for MPC via Deep Inverse Reinforcement Learning17 Jan 2022 0 repositories listed
-
State of the Art of Reinforcement Learning17 Jan 2022 0 repositories listed
-
Summarising and Comparing Agent Dynamics with Contrastive Spatiotemporal Abstraction17 Jan 2022 0 repositories listed
-
Towards deep observation: A systematic survey on artificial intelligence techniques to monitor fetus via Ultrasound Images17 Jan 2022 0 repositories listed
-
A Family of Cognitively Realistic Parsing Environments for Deep Reinforcement Learning16 Jan 2022 0 repositories listed
-
CONQRR: Conversational Query Rewriting for Retrieval with Reinforcement Learning16 Jan 2022 0 repositories listed
-
Inherently Explainable Reinforcement Learning in Natural Language16 Jan 2022 0 repositories listed
-
Jointly Reinforced User Simulator and Task-oriented Dialog System with Simplified Generative Architecture16 Jan 2022 0 repositories listed
-
Learning from Atypical Behavior: Temporary Interest Aware Recommendation Based on Reinforcement Learning16 Jan 2022 0 repositories listed
-
Parsing Natural Language into Propositional and First-Order Logic with Dual Reinforcement Learning16 Jan 2022 0 repositories listed
-
Reinforcement Learning with Large Action Spaces for Neural Machine Translation16 Jan 2022 0 repositories listed
-
Unsupervised Reinforcement Adaptation for Class-Imbalanced TextClassification16 Jan 2022 0 repositories listed
-
Block Policy Mirror Descent15 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning for Shared Autonomous Vehicles (SAV) Fleet Management15 Jan 2022 0 repositories listed
-
Interpretable and Effective Reinforcement Learning for Attacking against Graph-based Rumor Detection15 Jan 2022 0 repositories listed
-
Profitable Strategy Design by Using Deep Reinforcement Learning for Trades on Cryptocurrency Markets15 Jan 2022 0 repositories listed
-
Recursive Least Squares Advantage Actor-Critic Algorithms15 Jan 2022 0 repositories listed
-
Reinforcement Learning based Air Combat Maneuver Generation14 Jan 2022 0 repositories listed
-
Demystifying Reinforcement Learning in Time-Varying Systems14 Jan 2022 0 repositories listed
-
Reinforcement Learning to Solve NP-hard Problems: an Application to the CVRP14 Jan 2022 0 repositories listed
-
Automated Reinforcement Learning: An Overview13 Jan 2022 0 repositories listed
-
Criticality-Based Varying Step-Number Algorithm for Reinforcement Learning13 Jan 2022 0 repositories listed
-
Direct Mutation and Crossover in Genetic Algorithms Applied to Reinforcement Learning Tasks13 Jan 2022 0 repositories listed
-
The Recurrent Reinforcement Learning Crypto Agent12 Jan 2022 0 repositories listed
-
Active Reinforcement Learning -- A Roadmap Towards Curious Classifier Systems for Self-Adaptation11 Jan 2022 0 repositories listed
-
Automated Reinforcement Learning (AutoRL): A Survey and Open Problems11 Jan 2022 0 repositories listed
-
Benchmarking Deep Reinforcement Learning Algorithms for Vision-based Robotics11 Jan 2022 0 repositories listed
-
Pavlovian Signalling with General Value Functions in Agent-Agent Temporal Decision Making11 Jan 2022 0 repositories listed
-
Task Independent Capsule-Based Agents for Deep Q-Learning11 Jan 2022 0 repositories listed
-
Distributed Cooperative Multi-Agent Reinforcement Learning with Directed Coordination Graph10 Jan 2022 0 repositories listed
-
State of the Art of User Simulation approaches for conversational information retrieval10 Jan 2022 0 repositories listed
-
When is Offline Two-Player Zero-Sum Markov Game Solvable?10 Jan 2022 0 repositories listed
-
A Multi-agent Reinforcement Learning Approach for Efficient Client Selection in Federated Learning9 Jan 2022 0 repositories listed
-
Assessing Policy, Loss and Planning Combinations in Reinforcement Learning using a New Modular Architecture8 Jan 2022 0 repositories listed
-
Neural Network Optimization for Reinforcement Learning Tasks Using Sparse Computations7 Jan 2022 0 repositories listed
-
Offline Reinforcement Learning for Road Traffic Control7 Jan 2022 0 repositories listed
-
Combining Reinforcement Learning and Inverse Reinforcement Learning for Asset Allocation Recommendations6 Jan 2022 0 repositories listed
-
Offsetting Unequal Competition through RL-assisted Incentive Schemes5 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning, a textbook4 Jan 2022 0 repositories listed
-
A Deeper Understanding of State-Based Critics in Multi-Agent Reinforcement Learning3 Jan 2022 0 repositories listed
-
Actor-Critic Network for Q&A in an Adversarial Environment3 Jan 2022 0 repositories listed
-
Execute Order 66: Targeted Data Poisoning for Reinforcement Learning3 Jan 2022 0 repositories listed
-
Analyzing Micro-Founded General Equilibrium Models with Many Agents using Deep Reinforcement Learning3 Jan 2022 0 repositories listed
-
Reinforcement Learning for Task Specifications with Action-Constraints2 Jan 2022 0 repositories listed
-
Robust Algorithmic Collusion2 Jan 2022 0 repositories listed
-
A Surrogate-Assisted Controller for Expensive Evolutionary Reinforcement Learning1 Jan 2022 0 repositories listed
-
Operator Deep Q-Learning: Zero-Shot Reward Transferring in Reinforcement Learning1 Jan 2022 0 repositories listed
-
Temporal Complementarity-Guided Reinforcement Learning for Image-to-Video Person Re-Identification1 Jan 2022 0 repositories listed
-
Toward Pareto Efficient Fairness-Utility Trade-off inRecommendation through Reinforcement Learning1 Jan 2022 0 repositories listed
-
Importance of Empirical Sample Complexity Analysis for Offline Reinforcement Learning31 Dec 2021 0 repositories listed
-
Single-Shot Pruning for Offline Reinforcement Learning31 Dec 2021 0 repositories listed
-
Stochastic convex optimization for provably efficient apprenticeship learning31 Dec 2021 0 repositories listed
-
Using Graph-Aware Reinforcement Learning to Identify Winning Strategies in Diplomacy Games (Student Abstract)31 Dec 2021 0 repositories listed
-
Constructing a Good Behavior Basis for Transfer using Generalized Policy Updates30 Dec 2021 0 repositories listed
-
Multi-Agent Reinforcement Learning via Adaptive Kalman Temporal Difference and Successor Representation30 Dec 2021 0 repositories listed
-
Reversible Upper Confidence Bound Algorithm to Generate Diverse Optimized Candidates30 Dec 2021 0 repositories listed
-
Stability-Preserving Automatic Tuning of PID Control with Reinforcement Learning30 Dec 2021 0 repositories listed
-
Control Theoretic Analysis of Temporal Difference Learning29 Dec 2021 0 repositories listed
-
Efficient Performance Bounds for Primal-Dual Reinforcement Learning from Demonstrations28 Dec 2021 0 repositories listed
-
Robustness and risk management via distributional dynamic programming28 Dec 2021 0 repositories listed
-
Multiagent Model-based Credit Assignment for Continuous Control27 Dec 2021 0 repositories listed
-
RELDEC: Reinforcement Learning-Based Decoding of Moderate Length LDPC Codes27 Dec 2021 0 repositories listed
-
Safe Reinforcement Learning with Chance-constrained Model Predictive Control27 Dec 2021 0 repositories listed
-
The Statistical Complexity of Interactive Decision Making27 Dec 2021 0 repositories listed
-
Abstractions of General Reinforcement Learning26 Dec 2021 0 repositories listed
-
Neuro-Symbolic Hierarchical Rule Induction26 Dec 2021 0 repositories listed
-
Reducing Planning Complexity of General Reinforcement Learning with Non-Markovian Abstractions26 Dec 2021 0 repositories listed
-
A Survey on Interpretable Reinforcement Learning24 Dec 2021 0 repositories listed
-
Dynamic Channel Access via Meta-Reinforcement Learning24 Dec 2021 0 repositories listed
-
Rediscovering Affordance: A Reinforcement Learning Perspective24 Dec 2021 0 repositories listed
-
Improving the Efficiency of Off-Policy Reinforcement Learning by Accounting for Past Decisions23 Dec 2021 0 repositories listed
-
Local Advantage Networks for Cooperative Multi-Agent Reinforcement Learning23 Dec 2021 0 repositories listed
-
Missing Velocity in Dynamic Obstacle Avoidance based on Deep Reinforcement Learning23 Dec 2021 0 repositories listed
-
Deep Reinforcement Learning for Optimal Power Flow with Renewables Using Graph Information22 Dec 2021 0 repositories listed
-
Graph augmented Deep Reinforcement Learning in the GameRLand3D environment22 Dec 2021 0 repositories listed
-
Aerial Base Station Positioning and Power Control for Securing Communications: A Deep Q-Network Approach21 Dec 2021 0 repositories listed
-
Do Androids Dream of Electric Fences? Safety-Aware Reinforcement Learning with Latent Shielding21 Dec 2021 0 repositories listed
-
Reinforcement Learning based Sequential Batch-sampling for Bayesian Optimal Experimental Design21 Dec 2021 0 repositories listed
-
AGPNet -- Autonomous Grading Policy Network20 Dec 2021 0 repositories listed
-
Interpretable Preference-based Reinforcement Learning with Tree-Structured Reward Functions20 Dec 2021 0 repositories listed
-
CGIBNet: Bandwidth-constrained Communication with Graph Information Bottleneck in Multi-Agent Reinforcement Learning20 Dec 2021 0 repositories listed
-
Safe multi-agent deep reinforcement learning for joint bidding and maintenance scheduling of generation units20 Dec 2021 0 repositories listed
-
RoboAssembly: Learning Generalizable Furniture Assembly Policy in a Novel Multi-robot Contact-rich Simulation Environment19 Dec 2021 0 repositories listed
-
Creativity of AI: Hierarchical Planning Model Learning for Facilitating Deep Reinforcement Learning18 Dec 2021 0 repositories listed
-
Curriculum Based Reinforcement Learning of Grid Topology Controllers to Prevent Thermal Cascading18 Dec 2021 0 repositories listed
-
Contrastive Explanations for Comparing Preferences of Reinforcement Learning Agents17 Dec 2021 0 repositories listed
-
Deep Reinforcement Learning-based Authentic Dialogue Generation To Protect Youth From Cybergrooming17 Dec 2021 0 repositories listed
-
Learning Reward Machines: A Study in Partially Observable Reinforcement Learning17 Dec 2021 0 repositories listed
-
Personalized Lane Change Decision Algorithm Using Deep Reinforcement Learning Approach17 Dec 2021 0 repositories listed