Browse State-of-the-Art › reinforcement-learning › Papers, page 84
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 84 of 135: papers 8,301 to 8,400 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Sample-Efficient Reinforcement Learning with loglog(T) Switching Cost13 Feb 2022 0 repositories listed
-
End-to-end Reinforcement Learning of Robotic Manipulation with Robust Keypoints Representation12 Feb 2022 0 repositories listed
-
Neural NID Rules12 Feb 2022 0 repositories listed
-
Computational-Statistical Gaps in Reinforcement Learning11 Feb 2022 0 repositories listed
-
Rate-matching the regret lower-bound in the linear quadratic regulator with unknown dynamics11 Feb 2022 0 repositories listed
-
Regularized Q-learning11 Feb 2022 0 repositories listed
-
Abstraction for Deep Reinforcement Learning10 Feb 2022 0 repositories listed
-
AI-based Robust Resource Allocation in End-to-End Network Slicing under Demand and CSI Uncertainties10 Feb 2022 0 repositories listed
-
Group-Agent Reinforcement Learning10 Feb 2022 0 repositories listed
-
Interpretable pipelines with evolutionarily optimized modules for RL tasks with visual inputs10 Feb 2022 0 repositories listed
-
Reinforcement Learning in the Wild: Scalable RL Dispatching Algorithm Deployed in Ridehailing Marketplace10 Feb 2022 0 repositories listed
-
SAFER: Data-Efficient and Safe Reinforcement Learning via Skill Acquisition10 Feb 2022 0 repositories listed
-
Settling the Communication Complexity for Distributed Offline Reinforcement Learning10 Feb 2022 0 repositories listed
-
Understanding Value Decomposition Algorithms in Deep Cooperative Multi-Agent Reinforcement Learning10 Feb 2022 0 repositories listed
-
Universal Learning Waveform Selection Strategies for Adaptive Target Tracking10 Feb 2022 0 repositories listed
-
Offline Reinforcement Learning with Realizability and Single-policy Concentrability9 Feb 2022 0 repositories listed
-
Scenario-Assisted Deep Reinforcement Learning9 Feb 2022 0 repositories listed
-
PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning8 Feb 2022 0 repositories listed
-
Local Explanations for Reinforcement Learning8 Feb 2022 0 repositories listed
-
Provable Reinforcement Learning with a Short-Term Memory8 Feb 2022 0 repositories listed
-
Robust, Deep, and Reinforcement Learning for Management of Communication and Power Networks8 Feb 2022 0 repositories listed
-
Attacking c-MARL More Effectively: A Data Driven Approach7 Feb 2022 0 repositories listed
-
Model-Based Offline Meta-Reinforcement Learning with Regularization7 Feb 2022 0 repositories listed
-
Policy Optimization for Stochastic Shortest Path7 Feb 2022 0 repositories listed
-
Reward-Respecting Subtasks for Model-Based Reinforcement Learning7 Feb 2022 0 repositories listed
-
Exploration with Multi-Sample Target Values for Distributional Reinforcement Learning6 Feb 2022 0 repositories listed
-
Stochastic Gradient Descent with Dependent Data for Offline Reinforcement Learning6 Feb 2022 0 repositories listed
-
ASHA: Assistive Teleoperation via Human-in-the-Loop Reinforcement Learning5 Feb 2022 0 repositories listed
-
Meta-Reinforcement Learning with Self-Modifying Networks4 Feb 2022 0 repositories listed
-
A Reinforcement Learning Framework for PQoS in a Teleoperated Driving Scenario4 Feb 2022 0 repositories listed
-
Malleable Agents for Re-Configurable Robotic Manipulators4 Feb 2022 0 repositories listed
-
Model-Free Reinforcement Learning for Symbolic Automata-encoded Objectives4 Feb 2022 0 repositories listed
-
Offline Reinforcement Learning for Mobile Notifications4 Feb 2022 0 repositories listed
-
4 Feb 2022 0 repositories listed
-
AI-as-a-Service Toolkit for Human-Centered Intelligence in Autonomous Driving3 Feb 2022 0 repositories listed
-
Challenging Common Assumptions in Convex Reinforcement Learning3 Feb 2022 0 repositories listed
-
Deep Reinforcement Learning Assisted Federated Learning Algorithm for Data Management of IIoT3 Feb 2022 0 repositories listed
-
Financial Vision Based Reinforcement Learning Trading Strategy3 Feb 2022 0 repositories listed
-
How to Leverage Unlabeled Data in Offline Reinforcement Learning3 Feb 2022 0 repositories listed
-
Network Resource Allocation Strategy Based on Deep Reinforcement Learning3 Feb 2022 0 repositories listed
-
Reward is not enough: can we liberate AI from the reinforcement learning paradigm?3 Feb 2022 0 repositories listed
-
Security-Aware Virtual Network Embedding Algorithm based on Reinforcement Learning3 Feb 2022 0 repositories listed
-
Adaptive Discrete Communication Bottlenecks with Dynamic Vector Quantization2 Feb 2022 0 repositories listed
-
Federated Reinforcement Learning for Collective Navigation of Robotic Swarms2 Feb 2022 0 repositories listed
-
Transfer in Reinforcement Learning via Regret Bounds for Learning Agents2 Feb 2022 0 repositories listed
-
Improving Sample Efficiency of Value Based Models Using Attention and Vision Transformers1 Feb 2022 0 repositories listed
-
Reinforcement learning of optimal active particle navigation1 Feb 2022 0 repositories listed
-
Scalable Fragment-Based 3D Molecular Design with Reinforcement Learning1 Feb 2022 0 repositories listed
-
Sequential Search with Off-Policy Reinforcement Learning1 Feb 2022 0 repositories listed
-
Compositional Multi-Object Reinforcement Learning with Linear Relation Networks31 Jan 2022 0 repositories listed
-
Score vs. Winrate in Score-Based Games: which Reward for Reinforcement Learning?31 Jan 2022 0 repositories listed
-
Near-Optimal Regret for Adversarial MDP with Delayed Bandit Feedback31 Jan 2022 0 repositories listed
-
On solutions of the distributional Bellman equation31 Jan 2022 0 repositories listed
-
Reinforcement Learning with Heterogeneous Data: Estimation and Inference31 Jan 2022 0 repositories listed
-
Warmth and competence in human-agent cooperation31 Jan 2022 0 repositories listed
-
Communication-Efficient Consensus Mechanism for Federated Reinforcement Learning30 Jan 2022 0 repositories listed
-
Contrastive Learning from Demonstrations30 Jan 2022 0 repositories listed
-
Coordinated Frequency Control through Safe Reinforcement Learning30 Jan 2022 0 repositories listed
-
DearFSAC: An Approach to Optimizing Unreliable Federated Learning via Deep Reinforcement Learning30 Jan 2022 0 repositories listed
-
ApolloRL: a Reinforcement Learning Platform for Autonomous Driving29 Jan 2022 0 repositories listed
-
DeepRNG: Towards Deep Reinforcement Learning-Assisted Generative Testing of Software29 Jan 2022 0 repositories listed
-
A deep Q-learning method for optimizing visual search strategies in backgrounds of dynamic noise28 Jan 2022 0 repositories listed
-
Discovering Exfiltration Paths Using Reinforcement Learning with Attack Graphs28 Jan 2022 0 repositories listed
-
Dynamic Temporal Reconciliation by Reinforcement learning28 Jan 2022 0 repositories listed
-
Joint Differentiable Optimization and Verification for Certified Reinforcement Learning28 Jan 2022 0 repositories listed
-
Overcoming Exploration: Deep Reinforcement Learning for Continuous Control in Cluttered Environments from Temporal Logic Specifications28 Jan 2022 0 repositories listed
-
Generative Adversarial Exploration for Reinforcement Learning27 Jan 2022 0 repositories listed
-
Human-centered mechanism design with Democratic AI27 Jan 2022 0 repositories listed
-
Quantile-Based Policy Optimization for Reinforcement Learning27 Jan 2022 0 repositories listed
-
Multi-Agent Reinforcement Learning for Network Load Balancing in Data Center27 Jan 2022 0 repositories listed
-
Reinforcement Learning-Empowered Mobile Edge Computing for 6G Edge Intelligence27 Jan 2022 0 repositories listed
-
The Challenges of Exploration for Offline Reinforcement Learning27 Jan 2022 0 repositories listed
-
Exploiting Semantic Epsilon Greedy Exploration Strategy in Multi-Agent Reinforcement Learning26 Jan 2022 0 repositories listed
-
Hyperparameter Tuning for Deep Reinforcement Learning Applications26 Jan 2022 0 repositories listed
-
Probe-Based Interventions for Modifying Agent Behavior26 Jan 2022 0 repositories listed
-
MOORe: Model-based Offline-to-Online Reinforcement Learning25 Jan 2022 0 repositories listed
-
Reinforcement Learning Based Query Vertex Ordering Model for Subgraph Matching25 Jan 2022 0 repositories listed
-
Using Deep Reinforcement Learning for Zero Defect Smart Forging25 Jan 2022 0 repositories listed
-
Accelerated Intravascular Ultrasound Imaging using Deep Reinforcement Learning24 Jan 2022 0 repositories listed
-
State-Conditioned Adversarial Subgoal Generation24 Jan 2022 0 repositories listed
-
Large-Scale Graph Reinforcement Learning in Wireless Control Systems24 Jan 2022 0 repositories listed
-
Deep reinforcement learning under signal temporal logic constraints using Lagrangian relaxation21 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning with Spiking Q-learning21 Jan 2022 0 repositories listed
-
Instance-Dependent Confidence and Early Stopping for Reinforcement Learning21 Jan 2022 0 repositories listed
-
Learning Two-Step Hybrid Policy for Graph-Based Interpretable Reinforcement Learning21 Jan 2022 0 repositories listed
-
Reinforcement Learning for Personalized Drug Discovery and Design for Complex Diseases: A Systems Pharmacology Perspective21 Jan 2022 0 repositories listed
-
Reinforcement Learning Your Way: Agent Characterization through Policy Regularization21 Jan 2022 0 repositories listed
-
A Prescriptive Dirichlet Power Allocation Policy with Deep Reinforcement Learning20 Jan 2022 0 repositories listed
-
Learning Multi-agent Skills for Tabular Reinforcement Learning using Factor Graphs20 Jan 2022 0 repositories listed
-
Priors, Hierarchy, and Information Asymmetry for Skill Transfer in Reinforcement Learning20 Jan 2022 0 repositories listed
-
Recursive Constraints to Prevent Instability in Constrained Reinforcement Learning20 Jan 2022 0 repositories listed
-
Resource allocation algorithm for MEC based on Deep Reinforcement Learning20 Jan 2022 0 repositories listed
-
Safety-Aware Multi-Agent Apprenticeship Learning20 Jan 2022 0 repositories listed
-
Self-Awareness Safety of Deep Reinforcement Learning in Road Traffic Junction Driving20 Jan 2022 0 repositories listed
-
Sim-to-Lab-to-Real: Safe Reinforcement Learning with Shielding and Generalization Guarantees20 Jan 2022 0 repositories listed
-
Anytime PSRO for Two-Player Zero-Sum Games19 Jan 2022 0 repositories listed
-
Hybrid Reinforcement Learning-Based Eco-Driving Strategy for Connected and Automated Vehicles at Signalized Intersections19 Jan 2022 0 repositories listed
-
Online POI Recommendation: Learning Dynamic Geo-Human Interactions in Streams19 Jan 2022 0 repositories listed
-
Accelerating Representation Learning with View-Consistent Dynamics in Data-Efficient Reinforcement Learning18 Jan 2022 0 repositories listed
-
Conservative Distributional Reinforcement Learning with Safety Constraints18 Jan 2022 0 repositories listed