Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 97
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 97 of 152: papers 9,601 to 9,700 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A MultiModal Social Robot Toward Personalized Emotion Interaction8 Oct 2021 0 repositories listed
-
CheerBots: Chatbots toward Empathy and Emotionusing Reinforcement Learning8 Oct 2021 0 repositories listed
-
Learning to Centralize Dual-Arm Assembly8 Oct 2021 0 repositories listed
-
Revisiting Design Choices in Offline Model-Based Reinforcement Learning8 Oct 2021 0 repositories listed
-
Showing Your Offline Reinforcement Learning Work: Online Evaluation Budget Matters8 Oct 2021 0 repositories listed
-
A Model Selection Approach for Corruption Robust Reinforcement Learning7 Oct 2021 0 repositories listed
-
Bad-Policy Density: A Measure of Reinforcement Learning Hardness7 Oct 2021 0 repositories listed
-
Designing Composites with Target Effective Young's Modulus using Reinforcement Learning7 Oct 2021 0 repositories listed
-
Explaining Deep Reinforcement Learning Agents In The Atari Domain through a Surrogate Model7 Oct 2021 0 repositories listed
-
Generalization in Deep RL for TSP Problems via Equivariance and Local Search7 Oct 2021 0 repositories listed
-
Learning Pessimism for Robust and Efficient Off-Policy Reinforcement Learning7 Oct 2021 0 repositories listed
-
Near-Optimal Reward-Free Exploration for Linear Mixture MDPs with Plug-in Solver7 Oct 2021 0 repositories listed
-
Reinforcement Learning in Reward-Mixing MDPs7 Oct 2021 0 repositories listed
-
Robotic Lever Manipulation using Hindsight Experience Replay and Shapley Additive Explanations7 Oct 2021 0 repositories listed
-
The Benefits of Being Categorical Distributional: Uncertainty-aware Regularized Exploration in Reinforcement Learning7 Oct 2021 0 repositories listed
-
Adaptive control of a mechatronic system using constrained residual reinforcement learning6 Oct 2021 0 repositories listed
-
Scalable Multi-Agent Reinforcement Learning for Residential Load Scheduling under Data Governance6 Oct 2021 0 repositories listed
-
Heterogeneous Attentions for Solving Pickup and Delivery Problem via Deep Reinforcement Learning6 Oct 2021 0 repositories listed
-
Improving Generalization of Deep Reinforcement Learning-based TSP Solvers6 Oct 2021 0 repositories listed
-
Compositional Q-learning for electrolyte repletion with imbalanced patient sub-populations6 Oct 2021 0 repositories listed
-
Pretraining & Reinforcement Learning: Sharpening the Axe Before Cutting the Tree6 Oct 2021 0 repositories listed
-
A Deep Reinforcement Learning Framework for Contention-Based Spectrum Sharing5 Oct 2021 0 repositories listed
-
A study of first-passage time minimization via Q-learning in heated gridworlds5 Oct 2021 0 repositories listed
-
Decentralized Cooperative Lane Changing at Freeway Weaving Areas Using Multi-Agent Deep Reinforcement Learning5 Oct 2021 0 repositories listed
-
Deep reinforcement learning for guidewire navigation in coronary artery phantom5 Oct 2021 0 repositories listed
-
DeepEdge: A Deep Reinforcement Learning based Task Orchestrator for Edge Computing5 Oct 2021 0 repositories listed
-
NaRLE: Natural Language Models using Reinforcement Learning with Emotion Feedback5 Oct 2021 0 repositories listed
-
OTTR: Off-Road Trajectory Tracking using Reinforcement Learning5 Oct 2021 0 repositories listed
-
Mining for Potent Inhibitors through Artificial Intelligence and Physics: A Unified Methodology for Ligand Based and Structure Based Drug Design5 Oct 2021 0 repositories listed
-
You Only Evaluate Once: a Simple Baseline Algorithm for Offline RL5 Oct 2021 0 repositories listed
-
A Modified Q-Learning Algorithm for Rate-Profiling of Polarization Adjusted Convolutional (PAC) Codes4 Oct 2021 0 repositories listed
-
Automating Privilege Escalation with Deep Reinforcement Learning4 Oct 2021 0 repositories listed
-
Behaviour-conditioned policies for cooperative reinforcement learning tasks4 Oct 2021 0 repositories listed
-
4 Oct 2021 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 2 pointer-only (licence)
-
Learning to Assist Agents by Observing Them4 Oct 2021 0 repositories listed
-
Multi-Agent Path Planning Using Deep Reinforcement Learning4 Oct 2021 0 repositories listed
-
Reinforcement Learning for Admission Control in Wireless Virtual Network Embedding4 Oct 2021 0 repositories listed
-
A Novel Automated Curriculum Strategy to Solve Hard Sokoban Planning Instances3 Oct 2021 0 repositories listed
-
Decentralized Safe Reinforcement Learning for Voltage Control3 Oct 2021 0 repositories listed
-
DRL-Clusters: Buffer Management with Clustering based Deep Reinforcement Learning3 Oct 2021 0 repositories listed
-
Parallel Actors and Learners: A Framework for Generating Scalable RL Implementations3 Oct 2021 0 repositories listed
-
Feel-Good Thompson Sampling for Contextual Bandits and Reinforcement Learning2 Oct 2021 0 repositories listed
-
Seeking Visual Discomfort: Curiosity-driven Representations for Reinforcement Learning2 Oct 2021 0 repositories listed
-
Cellular traffic offloading via Opportunistic Networking with Reinforcement Learning1 Oct 2021 0 repositories listed
-
Divergence-Regularized Multi-Agent Actor-Critic1 Oct 2021 0 repositories listed
-
DNN-Opt: An RL Inspired Optimization for Analog Circuit Sizing using Deep Neural Networks1 Oct 2021 0 repositories listed
-
Motion Planning for Autonomous Vehicles in the Presence of Uncertainty Using Reinforcement Learning1 Oct 2021 0 repositories listed
-
Multi-lane Cruising Using Hierarchical Planning and Reinforcement Learning1 Oct 2021 0 repositories listed
-
Safety aware model-based reinforcement learning for optimal control of a class of output-feedback nonlinear systems1 Oct 2021 0 repositories listed
-
Terminal Adaptive Guidance for Autonomous Hypersonic Strike Weapons via Reinforcement Learning1 Oct 2021 0 repositories listed
-
A Privacy-preserving Distributed Training Framework for Cooperative Multi-agent Deep Reinforcement Learning30 Sep 2021 0 repositories listed
-
Bitcoin Transaction Strategy Construction Based on Deep Reinforcement Learning30 Sep 2021 0 repositories listed
-
Coordinated Reinforcement Learning for Optimizing Mobile Networks30 Sep 2021 0 repositories listed
-
Decentralized Graph-Based Multi-Agent Reinforcement Learning Using Reward Machines30 Sep 2021 0 repositories listed
-
HLIC: Harmonizing Optimization Metrics in Learned Image Compression by Reinforcement Learning30 Sep 2021 0 repositories listed
-
Modeling Interactions of Autonomous Vehicles and Pedestrians with Deep Multi-Agent Reinforcement Learning for Collision Avoidance30 Sep 2021 0 repositories listed
-
Reinforcement Learning for Classical Planning: Viewing Heuristics as Dense Reward Generators30 Sep 2021 0 repositories listed
-
Reinforcement Learning with Information-Theoretic Actuation30 Sep 2021 0 repositories listed
-
Stability Constrained Reinforcement Learning for Real-Time Voltage Control30 Sep 2021 0 repositories listed
-
Trajectory Planning with Deep Reinforcement Learning in High-Level Action Spaces30 Sep 2021 0 repositories listed
-
A Flexible Measurement of Diversity in Datasets with Random Network Distillation29 Sep 2021 0 repositories listed
-
A General Theory of Relativity in Reinforcement Learning29 Sep 2021 0 repositories listed
-
A Principled Permutation Invariant Approach to Mean-Field Multi-Agent Reinforcement Learning29 Sep 2021 0 repositories listed
-
A Two-Time-Scale Stochastic Optimization Framework with Applications in Control and Reinforcement Learning29 Sep 2021 0 repositories listed
-
AARL: Automated Auxiliary Loss for Reinforcement Learning29 Sep 2021 0 repositories listed
-
Adaptive Graph Capsule Convolutional Networks29 Sep 2021 0 repositories listed
-
Adaptive Q-learning for Interaction-Limited Reinforcement Learning29 Sep 2021 0 repositories listed
-
Adversarial Style Transfer for Robust Policy Optimization in Reinforcement Learning29 Sep 2021 0 repositories listed
-
An Attempt to Model Human Trust with Reinforcement Learning29 Sep 2021 0 repositories listed
-
An Experimental Design Perspective on Exploration in Reinforcement Learning29 Sep 2021 0 repositories listed
-
An Optics Controlling Environment and Reinforcement Learning Benchmarks29 Sep 2021 0 repositories listed
-
Assessing Deep Reinforcement Learning Policies via Natural Corruptions at the Edge of Imperceptibility29 Sep 2021 0 repositories listed
-
Auto-Encoding Inverse Reinforcement Learning29 Sep 2021 0 repositories listed
-
Bayesian Exploration for Lifelong Reinforcement Learning29 Sep 2021 0 repositories listed
-
Benchmarking Sample Selection Strategies for Batch Reinforcement Learning29 Sep 2021 0 repositories listed
-
Better state exploration using action sequence equivalence29 Sep 2021 0 repositories listed
-
Bi-linear Value Networks for Multi-goal Reinforcement Learning29 Sep 2021 0 repositories listed
-
Boosted Curriculum Reinforcement Learning29 Sep 2021 0 repositories listed
-
Can Reinforcement Learning Efficiently Find Stackelberg-Nash Equilibria in General-Sum Markov Games?29 Sep 2021 0 repositories listed
-
CausalDyna: Improving Generalization of Dyna-style Reinforcement Learning via Counterfactual-Based Data Augmentation29 Sep 2021 0 repositories listed
-
Closed-Loop Control of Additive Manufacturing via Reinforcement Learning29 Sep 2021 0 repositories listed
-
Combinatorial Reinforcement Learning Based Scheduling for DNN Execution on Edge29 Sep 2021 0 repositories listed
-
Conditional Value-at-Risk for Quantitative Trading: A Direct Reinforcement Learning Approach29 Sep 2021 0 repositories listed
-
Continuous Deep Q-Learning in Optimal Control Problems: Normalized Advantage Functions Analysis29 Sep 2021 0 repositories listed
-
Convergent and Efficient Deep Q Learning Algorithm29 Sep 2021 0 repositories listed
-
Coordinated Attacks Against Federated Learning: A Multi-Agent Reinforcement Learning Approach29 Sep 2021 0 repositories listed
-
CubeTR: Learning to Solve the Rubik's Cube using Transformers29 Sep 2021 0 repositories listed
-
Data Sharing without Rewards in Multi-Task Offline Reinforcement Learning29 Sep 2021 0 repositories listed
-
Decentralized Cooperative Multi-Agent Reinforcement Learning with Exploration29 Sep 2021 0 repositories listed
-
Decentralized Cross-Entropy Method for Model-Based Reinforcement Learning29 Sep 2021 0 repositories listed
-
Decoupling Strategy and Surface Realization for Task-oriented Dialogues29 Sep 2021 0 repositories listed
-
Deep Ensemble Policy Learning29 Sep 2021 0 repositories listed
-
Deep Inverse Reinforcement Learning via Adversarial One-Class Classification29 Sep 2021 0 repositories listed
-
Deep Learning of Intrinsically Motivated Options in the Arcade Learning Environment29 Sep 2021 0 repositories listed
-
Detecting Worst-case Corruptions via Loss Landscape Curvature in Deep Reinforcement Learning29 Sep 2021 0 repositories listed
-
DiBB: Distributing Black-Box Optimization29 Sep 2021 0 repositories listed
-
Disentangling Generalization in Reinforcement Learning29 Sep 2021 0 repositories listed
-
29 Sep 2021 0 repositories listed
-
Distributional Decision Transformer for Hindsight Information Matching29 Sep 2021 0 repositories listed
-
Distributional Perturbation for Efficient Exploration in Distributional Reinforcement Learning29 Sep 2021 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.