Browse State-of-the-Art › Reinforcement Learning › Papers, page 76
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 76 of 132: papers 7,501 to 7,600 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Distributional Reinforcement Learning for Scheduling of Chemical Production Processes1 Mar 2022 0 repositories listed
-
DreamingV2: Reinforcement Learning with Discrete World Models without Reconstruction1 Mar 2022 0 repositories listed
-
Weakly Supervised Disentangled Representation for Goal-conditioned Reinforcement Learning28 Feb 2022 0 repositories listed
-
Distributed Multi-Agent Reinforcement Learning Based on Graph-Induced Local Value Functions26 Feb 2022 0 repositories listed
-
Consolidated Adaptive T-soft Update for Deep Reinforcement Learning25 Feb 2022 0 repositories listed
-
Context-Hierarchy Inverse Reinforcement Learning25 Feb 2022 0 repositories listed
-
Learning Dynamic Mechanisms in Unknown Environments: A Reinforcement Learning Approach25 Feb 2022 0 repositories listed
-
Reachability analysis in stochastic directed graphs by reinforcement learning25 Feb 2022 0 repositories listed
-
Evolving-to-Learn Reinforcement Learning Tasks with Spiking Neural Networks24 Feb 2022 0 repositories listed
-
Consistent Dropout for Policy Gradient Reinforcement Learning23 Feb 2022 0 repositories listed
-
Reinforcement Learning in Practice: Opportunities and Challenges23 Feb 2022 0 repositories listed
-
Learning Relative Return Policies With Upside-Down Reinforcement Learning23 Feb 2022 0 repositories listed
-
Reinforcement Learning from Demonstrations by Novel Interactive Expert and Application to Automatic Berthing Control Systems for Unmanned Surface Vessel23 Feb 2022 0 repositories listed
-
Behaviour-Diverse Automatic Penetration Testing: A Curiosity-Driven Multi-Objective Deep Reinforcement Learning Approach22 Feb 2022 0 repositories listed
-
Continual Auxiliary Task Learning22 Feb 2022 0 repositories listed
-
Multi-fidelity reinforcement learning framework for shape optimization22 Feb 2022 0 repositories listed
-
Reward-Free Policy Space Compression for Reinforcement Learning22 Feb 2022 0 repositories listed
-
Autonomous Warehouse Robot using Deep Q-Learning21 Feb 2022 0 repositories listed
-
Hybrid Learning for Orchestrating Deep Learning Inference in Multi-user Edge-cloud Networks21 Feb 2022 0 repositories listed
-
Rule Mining over Knowledge Graphs via Reinforcement Learning21 Feb 2022 0 repositories listed
-
A Behavior Regularized Implicit Policy for Offline Reinforcement Learning19 Feb 2022 0 repositories listed
-
Multi-task Safe Reinforcement Learning for Navigating Intersections in Dense Traffic19 Feb 2022 0 repositories listed
-
Robust Reinforcement Learning as a Stackelberg Game via Adaptively-Regularized Adversarial Training19 Feb 2022 0 repositories listed
-
Who Are the Best Adopters? User Selection Model for Free Trial Item Promotion19 Feb 2022 0 repositories listed
-
Can Interpretable Reinforcement Learning Manage Prosperity Your Way?18 Feb 2022 0 repositories listed
-
A Survey of Explainable Reinforcement Learning17 Feb 2022 0 repositories listed
-
Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization17 Feb 2022 0 repositories listed
-
Retrieval-Augmented Reinforcement Learning17 Feb 2022 0 repositories listed
-
Robust Reinforcement Learning via Genetic Curriculum17 Feb 2022 0 repositories listed
-
Branching Reinforcement Learning16 Feb 2022 0 repositories listed
-
Domain Adaptive Fake News Detection via Reinforcement Learning16 Feb 2022 0 repositories listed
-
Interpretable Reinforcement Learning with Multilevel Subgoal Discovery15 Feb 2022 0 repositories listed
-
Learning to Mitigate AI Collusion on Economic Platforms15 Feb 2022 0 repositories listed
-
User-Oriented Robust Reinforcement Learning15 Feb 2022 0 repositories listed
-
Provably Efficient Causal Model-Based Reinforcement Learning for Systematic Generalization14 Feb 2022 0 repositories listed
-
Reinforcement Learning in Presence of Discrete Markovian Context Evolution14 Feb 2022 0 repositories listed
-
Sequential Bayesian experimental designs via reinforcement learning14 Feb 2022 0 repositories listed
-
Statistical Inference After Adaptive Sampling for Longitudinal Data14 Feb 2022 0 repositories listed
-
Towards Deployment-Efficient Reinforcement Learning: Lower Bound and Optimality14 Feb 2022 0 repositories listed
-
Individual-Level Inverse Reinforcement Learning for Mean Field Games13 Feb 2022 0 repositories listed
-
Sample-Efficient Reinforcement Learning with loglog(T) Switching Cost13 Feb 2022 0 repositories listed
-
Computational-Statistical Gaps in Reinforcement Learning11 Feb 2022 0 repositories listed
-
Rate-matching the regret lower-bound in the linear quadratic regulator with unknown dynamics11 Feb 2022 0 repositories listed
-
Abstraction for Deep Reinforcement Learning10 Feb 2022 0 repositories listed
-
AI-based Robust Resource Allocation in End-to-End Network Slicing under Demand and CSI Uncertainties10 Feb 2022 0 repositories listed
-
Group-Agent Reinforcement Learning10 Feb 2022 0 repositories listed
-
Interpretable pipelines with evolutionarily optimized modules for RL tasks with visual inputs10 Feb 2022 0 repositories listed
-
SAFER: Data-Efficient and Safe Reinforcement Learning via Skill Acquisition10 Feb 2022 0 repositories listed
-
Settling the Communication Complexity for Distributed Offline Reinforcement Learning10 Feb 2022 0 repositories listed
-
Offline Reinforcement Learning with Realizability and Single-policy Concentrability9 Feb 2022 0 repositories listed
-
Scenario-Assisted Deep Reinforcement Learning9 Feb 2022 0 repositories listed
-
PolicyCleanse: Backdoor Detection and Mitigation in Reinforcement Learning8 Feb 2022 0 repositories listed
-
Local Explanations for Reinforcement Learning8 Feb 2022 0 repositories listed
-
Provable Reinforcement Learning with a Short-Term Memory8 Feb 2022 0 repositories listed
-
Model-Based Offline Meta-Reinforcement Learning with Regularization7 Feb 2022 0 repositories listed
-
Policy Optimization for Stochastic Shortest Path7 Feb 2022 0 repositories listed
-
Reward-Respecting Subtasks for Model-Based Reinforcement Learning7 Feb 2022 0 repositories listed
-
Exploration with Multi-Sample Target Values for Distributional Reinforcement Learning6 Feb 2022 0 repositories listed
-
Stochastic Gradient Descent with Dependent Data for Offline Reinforcement Learning6 Feb 2022 0 repositories listed
-
ASHA: Assistive Teleoperation via Human-in-the-Loop Reinforcement Learning5 Feb 2022 0 repositories listed
-
Meta-Reinforcement Learning with Self-Modifying Networks4 Feb 2022 0 repositories listed
-
Model-Free Reinforcement Learning for Symbolic Automata-encoded Objectives4 Feb 2022 0 repositories listed
-
Offline Reinforcement Learning for Mobile Notifications4 Feb 2022 0 repositories listed
-
4 Feb 2022 0 repositories listed
-
Challenging Common Assumptions in Convex Reinforcement Learning3 Feb 2022 0 repositories listed
-
Financial Vision Based Reinforcement Learning Trading Strategy3 Feb 2022 0 repositories listed
-
How to Leverage Unlabeled Data in Offline Reinforcement Learning3 Feb 2022 0 repositories listed
-
Network Resource Allocation Strategy Based on Deep Reinforcement Learning3 Feb 2022 0 repositories listed
-
Adaptive Discrete Communication Bottlenecks with Dynamic Vector Quantization2 Feb 2022 0 repositories listed
-
Federated Reinforcement Learning for Collective Navigation of Robotic Swarms2 Feb 2022 0 repositories listed
-
Transfer in Reinforcement Learning via Regret Bounds for Learning Agents2 Feb 2022 0 repositories listed
-
Improving Sample Efficiency of Value Based Models Using Attention and Vision Transformers1 Feb 2022 0 repositories listed
-
Reinforcement learning of optimal active particle navigation1 Feb 2022 0 repositories listed
-
Scalable Fragment-Based 3D Molecular Design with Reinforcement Learning1 Feb 2022 0 repositories listed
-
Sequential Search with Off-Policy Reinforcement Learning1 Feb 2022 0 repositories listed
-
Compositional Multi-Object Reinforcement Learning with Linear Relation Networks31 Jan 2022 0 repositories listed
-
On solutions of the distributional Bellman equation31 Jan 2022 0 repositories listed
-
Reinforcement Learning with Heterogeneous Data: Estimation and Inference31 Jan 2022 0 repositories listed
-
Warmth and competence in human-agent cooperation31 Jan 2022 0 repositories listed
-
Communication-Efficient Consensus Mechanism for Federated Reinforcement Learning30 Jan 2022 0 repositories listed
-
Coordinated Frequency Control through Safe Reinforcement Learning30 Jan 2022 0 repositories listed
-
ApolloRL: a Reinforcement Learning Platform for Autonomous Driving29 Jan 2022 0 repositories listed
-
DeepRNG: Towards Deep Reinforcement Learning-Assisted Generative Testing of Software29 Jan 2022 0 repositories listed
-
A deep Q-learning method for optimizing visual search strategies in backgrounds of dynamic noise28 Jan 2022 0 repositories listed
-
Discovering Exfiltration Paths Using Reinforcement Learning with Attack Graphs28 Jan 2022 0 repositories listed
-
Dynamic Temporal Reconciliation by Reinforcement learning28 Jan 2022 0 repositories listed
-
Joint Differentiable Optimization and Verification for Certified Reinforcement Learning28 Jan 2022 0 repositories listed
-
Generative Adversarial Exploration for Reinforcement Learning27 Jan 2022 0 repositories listed
-
Quantile-Based Policy Optimization for Reinforcement Learning27 Jan 2022 0 repositories listed
-
Reinforcement Learning-Empowered Mobile Edge Computing for 6G Edge Intelligence27 Jan 2022 0 repositories listed
-
The Challenges of Exploration for Offline Reinforcement Learning27 Jan 2022 0 repositories listed
-
Hyperparameter Tuning for Deep Reinforcement Learning Applications26 Jan 2022 0 repositories listed
-
MOORe: Model-based Offline-to-Online Reinforcement Learning25 Jan 2022 0 repositories listed
-
Reinforcement Learning Based Query Vertex Ordering Model for Subgraph Matching25 Jan 2022 0 repositories listed
-
Using Deep Reinforcement Learning for Zero Defect Smart Forging25 Jan 2022 0 repositories listed
-
Accelerated Intravascular Ultrasound Imaging using Deep Reinforcement Learning24 Jan 2022 0 repositories listed
-
Large-Scale Graph Reinforcement Learning in Wireless Control Systems24 Jan 2022 0 repositories listed
-
Deep Reinforcement Learning with Spiking Q-learning21 Jan 2022 0 repositories listed
-
Instance-Dependent Confidence and Early Stopping for Reinforcement Learning21 Jan 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.