Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 90
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 90 of 152: papers 8,901 to 9,000 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Domain Knowledge-Based Automated Analog Circuit Design with Deep Reinforcement Learning26 Feb 2022 0 repositories listed
-
Whittle Index based Q-Learning for Wireless Edge Caching with Linear Function Approximation26 Feb 2022 0 repositories listed
-
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning25 Feb 2022 0 repositories listed
-
Context-Hierarchy Inverse Reinforcement Learning25 Feb 2022 0 repositories listed
-
Decision Making in Non-Stationary Environments with Policy-Augmented Monte Carlo Tree Search25 Feb 2022 0 repositories listed
-
Learning Dynamic Mechanisms in Unknown Environments: A Reinforcement Learning Approach25 Feb 2022 0 repositories listed
-
Reachability analysis in stochastic directed graphs by reinforcement learning25 Feb 2022 0 repositories listed
-
Evolutionary Multi-Objective Reinforcement Learning Based Trajectory Control and Task Offloading in UAV-Assisted Mobile Edge Computing24 Feb 2022 0 repositories listed
-
Evolving-to-Learn Reinforcement Learning Tasks with Spiking Neural Networks24 Feb 2022 0 repositories listed
-
Comparative analysis of machine learning methods for active flow control23 Feb 2022 0 repositories listed
-
Consistent Dropout for Policy Gradient Reinforcement Learning23 Feb 2022 0 repositories listed
-
Reinforcement Learning in Practice: Opportunities and Challenges23 Feb 2022 0 repositories listed
-
Drawing Inductor Layout with a Reinforcement Learning Agent: Method and Application for VCO Inductors23 Feb 2022 0 repositories listed
-
Learning Relative Return Policies With Upside-Down Reinforcement Learning23 Feb 2022 0 repositories listed
-
Reinforcement Learning from Demonstrations by Novel Interactive Expert and Application to Automatic Berthing Control Systems for Unmanned Surface Vessel23 Feb 2022 0 repositories listed
-
Training Characteristic Functions with Reinforcement Learning: XAI-methods play Connect Four23 Feb 2022 0 repositories listed
-
A Decentralized Communication Framework based on Dual-Level Recurrence for Multi-Agent Reinforcement Learning22 Feb 2022 0 repositories listed
-
A policy gradient approach for optimization of smooth risk measures22 Feb 2022 0 repositories listed
-
Behaviour-Diverse Automatic Penetration Testing: A Curiosity-Driven Multi-Objective Deep Reinforcement Learning Approach22 Feb 2022 0 repositories listed
-
Behaviour-neutral Smart Charging of Plugin Electric Vehicles: Reinforcement learning approach22 Feb 2022 0 repositories listed
-
Continual Auxiliary Task Learning22 Feb 2022 0 repositories listed
-
Multi-fidelity reinforcement learning framework for shape optimization22 Feb 2022 0 repositories listed
-
Reward-Free Policy Space Compression for Reinforcement Learning22 Feb 2022 0 repositories listed
-
Sequential Information Design: Markov Persuasion Process and Its Efficient Reinforcement Learning22 Feb 2022 0 repositories listed
-
Accelerating Primal-dual Methods for Regularized Markov Decision Processes21 Feb 2022 0 repositories listed
-
Autonomous Warehouse Robot using Deep Q-Learning21 Feb 2022 0 repositories listed
-
CCPT: Automatic Gameplay Testing and Validation with Curiosity-Conditioned Proximal Trajectories21 Feb 2022 0 repositories listed
-
Hybrid Learning for Orchestrating Deep Learning Inference in Multi-user Edge-cloud Networks21 Feb 2022 0 repositories listed
-
Learning Causal Overhypotheses through Exploration in Children and Computational Models21 Feb 2022 0 repositories listed
-
Reinforcement Learning Framework for Server Placement and Workload Allocation in Multi-Access Edge Computing21 Feb 2022 0 repositories listed
-
Rule Mining over Knowledge Graphs via Reinforcement Learning21 Feb 2022 0 repositories listed
-
PooL: Pheromone-inspired Communication Framework forLarge Scale Multi-Agent Reinforcement Learning20 Feb 2022 0 repositories listed
-
Selective Credit Assignment20 Feb 2022 0 repositories listed
-
A Behavior Regularized Implicit Policy for Offline Reinforcement Learning19 Feb 2022 0 repositories listed
-
Multi-task Safe Reinforcement Learning for Navigating Intersections in Dense Traffic19 Feb 2022 0 repositories listed
-
Robust Reinforcement Learning as a Stackelberg Game via Adaptively-Regularized Adversarial Training19 Feb 2022 0 repositories listed
-
Transformation Coding: Simple Objectives for Equivariant Representations19 Feb 2022 0 repositories listed
-
Who Are the Best Adopters? User Selection Model for Free Trial Item Promotion19 Feb 2022 0 repositories listed
-
Can Interpretable Reinforcement Learning Manage Prosperity Your Way?18 Feb 2022 0 repositories listed
-
tinyMAN: Lightweight Energy Manager using Reinforcement Learning for Energy Harvesting Wearable IoT Devices18 Feb 2022 0 repositories listed
-
A Survey of Explainable Reinforcement Learning17 Feb 2022 0 repositories listed
-
A Survey on Deep Reinforcement Learning-based Approaches for Adaptation and Generalization17 Feb 2022 0 repositories listed
-
BADDr: Bayes-Adaptive Deep Dropout RL for POMDPs17 Feb 2022 0 repositories listed
-
Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization17 Feb 2022 0 repositories listed
-
Retrieval-Augmented Reinforcement Learning17 Feb 2022 0 repositories listed
-
Robust Reinforcement Learning via Genetic Curriculum17 Feb 2022 0 repositories listed
-
Should I send this notification? Optimizing push notifications decision making by modeling the future17 Feb 2022 0 repositories listed
-
UAV Base Station Trajectory Optimization Based on Reinforcement Learning in Post-disaster Search and Rescue Operations17 Feb 2022 0 repositories listed
-
Branching Reinforcement Learning16 Feb 2022 0 repositories listed
-
Domain Adaptive Fake News Detection via Reinforcement Learning16 Feb 2022 0 repositories listed
-
Policy Learning and Evaluation with Randomized Quasi-Monte Carlo16 Feb 2022 0 repositories listed
-
Deep Reinforcement Learning Based Multi-Access Edge Computing Schedule for Internet of Vehicle15 Feb 2022 0 repositories listed
-
Interpretable Reinforcement Learning with Multilevel Subgoal Discovery15 Feb 2022 0 repositories listed
-
L2C2: Locally Lipschitz Continuous Constraint towards Stable and Smooth Reinforcement Learning15 Feb 2022 0 repositories listed
-
Learning to Mitigate AI Collusion on Economic Platforms15 Feb 2022 0 repositories listed
-
User-Oriented Robust Reinforcement Learning15 Feb 2022 0 repositories listed
-
Convex Programs and Lyapunov Functions for Reinforcement Learning: A Unified Perspective on the Analysis of Value-Based Methods14 Feb 2022 0 repositories listed
-
Motivating Physical Activity via Competitive Human-Robot Interaction14 Feb 2022 0 repositories listed
-
Provably Efficient Causal Model-Based Reinforcement Learning for Systematic Generalization14 Feb 2022 0 repositories listed
-
Reinforcement Learning in Presence of Discrete Markovian Context Evolution14 Feb 2022 0 repositories listed
-
Robust Policy Learning over Multiple Uncertainty Sets14 Feb 2022 0 repositories listed
-
Sequential Bayesian experimental designs via reinforcement learning14 Feb 2022 0 repositories listed
-
Statistical Inference After Adaptive Sampling for Longitudinal Data14 Feb 2022 0 repositories listed
-
Towards Deployment-Efficient Reinforcement Learning: Lower Bound and Optimality14 Feb 2022 0 repositories listed
-
Autonomous Drone Swarm Navigation and Multi-target Tracking in 3D Environments with Dynamic Obstacles13 Feb 2022 0 repositories listed
-
Deep Reinforcement Learning and Convex Mean-Variance Optimisation for Portfolio Management13 Feb 2022 0 repositories listed
-
Individual-Level Inverse Reinforcement Learning for Mean Field Games13 Feb 2022 0 repositories listed
-
Sample-Efficient Reinforcement Learning with loglog(T) Switching Cost13 Feb 2022 0 repositories listed
-
End-to-end Reinforcement Learning of Robotic Manipulation with Robust Keypoints Representation12 Feb 2022 0 repositories listed
-
Neural NID Rules12 Feb 2022 0 repositories listed
-
A Unified Perspective on Value Backup and Exploration in Monte-Carlo Tree Search11 Feb 2022 0 repositories listed
-
Computational-Statistical Gaps in Reinforcement Learning11 Feb 2022 0 repositories listed
-
Rate-matching the regret lower-bound in the linear quadratic regulator with unknown dynamics11 Feb 2022 0 repositories listed
-
Regularized Q-learning11 Feb 2022 0 repositories listed
-
Abstraction for Deep Reinforcement Learning10 Feb 2022 0 repositories listed
-
AI-based Robust Resource Allocation in End-to-End Network Slicing under Demand and CSI Uncertainties10 Feb 2022 0 repositories listed
-
Group-Agent Reinforcement Learning10 Feb 2022 0 repositories listed
-
Interpretable pipelines with evolutionarily optimized modules for RL tasks with visual inputs10 Feb 2022 0 repositories listed
-
Off-Policy Fitted Q-Evaluation with Differentiable Function Approximators: Z-Estimation and Inference Theory10 Feb 2022 0 repositories listed
-
Reinforcement Learning in the Wild: Scalable RL Dispatching Algorithm Deployed in Ridehailing Marketplace10 Feb 2022 0 repositories listed
-
SAFER: Data-Efficient and Safe Reinforcement Learning via Skill Acquisition10 Feb 2022 0 repositories listed
-
Settling the Communication Complexity for Distributed Offline Reinforcement Learning10 Feb 2022 0 repositories listed
-
Understanding Value Decomposition Algorithms in Deep Cooperative Multi-Agent Reinforcement Learning10 Feb 2022 0 repositories listed
-
Universal Learning Waveform Selection Strategies for Adaptive Target Tracking10 Feb 2022 0 repositories listed
-
Intelligent Autonomous Intersection Management9 Feb 2022 0 repositories listed
-
Offline Reinforcement Learning with Realizability and Single-policy Concentrability9 Feb 2022 0 repositories listed
-
Scenario-Assisted Deep Reinforcement Learning9 Feb 2022 0 repositories listed
-
Transferred Q-learning9 Feb 2022 0 repositories listed
-
Understanding and Shifting Preferences for Battery Electric Vehicles9 Feb 2022 0 repositories listed
-
PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning8 Feb 2022 0 repositories listed
-
Energy Management Based on Multi-Agent Deep Reinforcement Learning for A Multi-Energy Industrial Park8 Feb 2022 0 repositories listed
-
GrASP: Gradient-Based Affordance Selection for Planning8 Feb 2022 0 repositories listed
-
Independent Policy Gradient for Large-Scale Markov Potential Games: Sharper Rates, Function Approximation, and Game-Agnostic Convergence8 Feb 2022 0 repositories listed
-
Local Explanations for Reinforcement Learning8 Feb 2022 0 repositories listed
-
Provable Reinforcement Learning with a Short-Term Memory8 Feb 2022 0 repositories listed
-
Robust, Deep, and Reinforcement Learning for Management of Communication and Power Networks8 Feb 2022 0 repositories listed
-
Attacking c-MARL More Effectively: A Data Driven Approach7 Feb 2022 0 repositories listed
-
Model-Based Offline Meta-Reinforcement Learning with Regularization7 Feb 2022 0 repositories listed
-
Policy Optimization for Stochastic Shortest Path7 Feb 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.