Browse State-of-the-Art › reinforcement-learning › Papers, page 67
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 67 of 135: papers 6,601 to 6,700 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Learning UI-to-Code Reverse Generator Using Visual Critic Without Rendering24 May 2023 0 repositories listed
-
Replicable Reinforcement Learning24 May 2023 0 repositories listed
-
Successor-Predecessor Intrinsic Exploration24 May 2023 0 repositories listed
-
Combining Multi-Objective Bayesian Optimization with Reinforcement Learning for TinyML23 May 2023 0 repositories listed
-
ChemGymRL: An Interactive Framework for Reinforcement Learning for Digital Chemistry23 May 2023 0 repositories listed
-
Constrained Proximal Policy Optimization23 May 2023 0 repositories listed
-
Control of a simulated MRI scanner with deep reinforcement learning23 May 2023 0 repositories listed
-
L-SA: Learning Under-Explored Targets in Multi-Target Reinforcement Learning23 May 2023 0 repositories listed
-
Language Model Self-improvement by Reinforcement Learning Contemplation23 May 2023 0 repositories listed
-
OER: Offline Experience Replay for Continual Offline Reinforcement Learning23 May 2023 0 repositories listed
-
Optimizing Long-term Value for Auction-Based Recommender Systems via On-Policy Reinforcement Learning23 May 2023 0 repositories listed
-
Proximal Policy Gradient Arborescence for Quality Diversity Reinforcement Learning23 May 2023 0 repositories listed
-
RLBoost: Boosting Supervised Models using Deep Reinforcement Learning23 May 2023 0 repositories listed
-
Semantic-aware Transmission Scheduling: a Monotonicity-driven Deep Reinforcement Learning Approach23 May 2023 0 repositories listed
-
Adaptive action supervision in reinforcement learning from real-world multi-agent demonstrations22 May 2023 0 repositories listed
-
Achieving the Asymptotically Optimal Sample Complexity of Offline Reinforcement Learning: A DRO-Based Approach22 May 2023 0 repositories listed
-
Lagrangian-based online safe reinforcement learning for state-constrained systems22 May 2023 0 repositories listed
-
Offline Primal-Dual Reinforcement Learning for Linear MDPs22 May 2023 0 repositories listed
-
Offline Reinforcement Learning with Additional Covering Distributions22 May 2023 0 repositories listed
-
TOM: Learning Policy-Aware Models for Model-Based Reinforcement Learning via Transition Occupancy Matching22 May 2023 0 repositories listed
-
A Reinforcement Learning Approach for Robust Supervisory Control of UAVs Under Disturbances21 May 2023 0 repositories listed
-
Towards Optimal Energy Management Strategy for Hybrid Electric Vehicle with Reinforcement Learning21 May 2023 0 repositories listed
-
A Framework for Provably Stable and Consistent Training of Deep Feedforward Networks20 May 2023 0 repositories listed
-
Game-Theoretical Analysis of Reviewer Rewards in Peer-Review Journal Systems: Analysis and Experimental Evaluation using Deep Reinforcement Learning20 May 2023 0 repositories listed
-
Model-based adaptation for sample efficient transfer in reinforcement learning control of parameter-varying systems20 May 2023 0 repositories listed
-
On First-Order Meta-Reinforcement Learning with Moreau Envelopes20 May 2023 0 repositories listed
-
Shattering the Agent-Environment Interface for Fine-Tuning Inclusive Language Models19 May 2023 0 repositories listed
-
Understanding the World to Solve Social Dilemmas Using Multi-Agent Reinforcement Learning19 May 2023 0 repositories listed
-
Automatic Design Method of Building Pipeline Layout Based on Deep Reinforcement Learning18 May 2023 0 repositories listed
-
Bayesian Reparameterization of Reward-Conditioned Reinforcement Learning with Energy-based Models18 May 2023 0 repositories listed
-
Black-Box Targeted Reward Poisoning Attack Against Online Deep Reinforcement Learning18 May 2023 0 repositories listed
-
Deep Metric Tensor Regularized Policy Gradient18 May 2023 0 repositories listed
-
Deep PackGen: A Deep Reinforcement Learning Framework for Adversarial Network Packet Generation18 May 2023 0 repositories listed
-
Parallel development of social preferences in fish and machines18 May 2023 0 repositories listed
-
Semantically Aligned Task Decomposition in Multi-Agent Reinforcement Learning18 May 2023 0 repositories listed
-
A proof of imitation of Wasserstein inverse reinforcement learning for multi-objective optimization17 May 2023 0 repositories listed
-
Curriculum Learning in Job Shop Scheduling using Reinforcement Learning17 May 2023 0 repositories listed
-
Discovering Individual Rewards in Collective Behavior through Inverse Multi-Agent Reinforcement Learning17 May 2023 0 repositories listed
-
Integrated Conflict Management for UAM with Strategic Demand Capacity Balancing and Learning-based Tactical Deconfliction17 May 2023 0 repositories listed
-
Model-Free Robust Average-Reward Reinforcement Learning17 May 2023 0 repositories listed
-
Multi-Agent Reinforcement Learning: Methods, Applications, Visionary Prospects, and Challenges17 May 2023 0 repositories listed
-
Pragmatic Reasoning in Structured Signaling Games17 May 2023 0 repositories listed
-
Reward-agnostic Fine-tuning: Provable Statistical Benefits of Hybrid Reinforcement Learning17 May 2023 0 repositories listed
-
Deep Reinforcement Learning to Maximize Arterial Usage during Extreme Congestion16 May 2023 0 repositories listed
-
Reinforcement Learning for Safe Robot Control using Control Lyapunov Barrier Functions16 May 2023 0 repositories listed
-
Horizon-free Reinforcement Learning in Adversarial Linear Mixture MDPs15 May 2023 0 repositories listed
-
Toward Multi-Agent Reinforcement Learning for Distributed Event-Triggered Control15 May 2023 0 repositories listed
-
Federated TD Learning over Finite-Rate Erasure Channels: Linear Speedup under Markovian Sampling14 May 2023 0 repositories listed
-
Inverse Reinforcement Learning With Constraint Recovery14 May 2023 0 repositories listed
-
Gaussian Prior Reinforcement Learning for Nested Named Entity Recognition12 May 2023 0 repositories listed
-
Identify, Estimate and Bound the Uncertainty of Reinforcement Learning for Autonomous Driving12 May 2023 0 repositories listed
-
Multi-Agent Reinforcement Learning for Network Routing in Integrated Access Backhaul Networks12 May 2023 0 repositories listed
-
S-REINFORCE: A Neuro-Symbolic Policy Gradient Approach for Interpretable Reinforcement Learning12 May 2023 0 repositories listed
-
12 May 2023 0 repositories listed Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample) · 1 pointer-only (licence)
-
Deep Reinforcement Learning for Interference Management in UAV-based 3D Networks: Potentials and Challenges11 May 2023 0 repositories listed
-
On Practical Robust Reinforcement Learning: Practical Uncertainty Set and Double-Agent Algorithm11 May 2023 0 repositories listed
-
Optimizing Memory Mapping Using Deep Reinforcement Learning11 May 2023 0 repositories listed
-
Towards Theoretical Understanding of Data-Driven Policy Refinement11 May 2023 0 repositories listed
-
A proof of convergence of inverse reinforcement learning for multi-objective optimization10 May 2023 0 repositories listed
-
An Option-Dependent Analysis of Regret Minimization Algorithms in Finite-Horizon Semi-Markov Decision Processes10 May 2023 0 repositories listed
-
Deep Reinforcement Learning Based Resource Allocation for Cloud Native Wireless Network10 May 2023 0 repositories listed
-
Discovery of Optimal Quantum Error Correcting Codes via Reinforcement Learning10 May 2023 0 repositories listed
-
HoneyIoT: Adaptive High-Interaction Honeypot for IoT Devices Through Reinforcement Learning10 May 2023 0 repositories listed
-
Mixture of personality improved Spiking actor network for efficient multi-agent cooperation10 May 2023 0 repositories listed
-
Cooperative Multi-Agent Reinforcement Learning: Asynchronous Communication and Linear Function Approximation10 May 2023 0 repositories listed
-
Supplementing Gradient-Based Reinforcement Learning with Simple Evolutionary Ideas10 May 2023 0 repositories listed
-
Assessment of Reinforcement Learning Algorithms for Nuclear Power Plant Fuel Optimization9 May 2023 0 repositories listed
-
Cooperating Graph Neural Networks with Deep Reinforcement Learning for Vaccine Prioritization9 May 2023 0 repositories listed
-
Fine-tuning Language Models with Generative Adversarial Reward Modelling9 May 2023 0 repositories listed
-
RLocator: Reinforcement Learning for Bug Localization9 May 2023 0 repositories listed
-
Adaptive Learning Path Navigation Based on Knowledge Tracing and Reinforcement Learning8 May 2023 0 repositories listed
-
Goal-oriented inference of environment from redundant observations8 May 2023 0 repositories listed
-
Truncating Trajectories in Monte Carlo Reinforcement Learning7 May 2023 0 repositories listed
-
A Survey on Offline Model-Based Reinforcement Learning5 May 2023 0 repositories listed
-
Bayesian Reinforcement Learning with Limited Cognitive Load5 May 2023 0 repositories listed
-
Reinforcement Learning for Control of Evolutionary and Ecological Processes5 May 2023 0 repositories listed
-
Improving Real-Time Bidding in Online Advertising Using Markov Decision Processes and Machine Learning Techniques5 May 2023 0 repositories listed
-
Knowledge Transfer from Teachers to Learners in Growing-Batch Reinforcement Learning5 May 2023 0 repositories listed
-
Maximum Causal Entropy Inverse Constrained Reinforcement Learning4 May 2023 0 repositories listed
-
Reinforcement Learning with Delayed, Composite, and Partially Anonymous Reward4 May 2023 0 repositories listed
-
Rethinking Population-assisted Off-policy Reinforcement Learning4 May 2023 0 repositories listed
-
Gym-preCICE: Reinforcement Learning Environments for Active Flow Control3 May 2023 0 repositories listed
-
Human Machine Co-adaption Interface via Cooperation Markov Decision Process System3 May 2023 0 repositories listed
-
An Improved Yaw Control Algorithm for Wind Turbines via Reinforcement Learning2 May 2023 0 repositories listed
-
Representations and Exploration for Deep Reinforcement Learning using Singular Value Decomposition1 May 2023 0 repositories listed
-
Joint Learning of Policy with Unknown Temporal Constraints for Safe Reinforcement Learning30 Apr 2023 0 repositories listed
-
A Transfer Learning Approach to Minimize Reinforcement Learning Risks in Energy Optimization for Smart Buildings30 Apr 2023 0 repositories listed
-
SRL-Assisted AFM: Generating Planar Unstructured Quadrilateral Meshes with Supervised and Reinforcement Learning-Assisted Advancing Front Method30 Apr 2023 0 repositories listed
-
Meta-Reinforcement Learning Based on Self-Supervised Task Representation Learning29 Apr 2023 0 repositories listed
-
Systematic Review on Reinforcement Learning in the Field of Fintech29 Apr 2023 0 repositories listed
-
A Federated Reinforcement Learning Framework for Link Activation in Multi-link Wi-Fi Networks28 Apr 2023 0 repositories listed
-
Active Reinforcement Learning for Personalized Stress Monitoring in Everyday Settings28 Apr 2023 0 repositories listed
-
Adversarial Policy Optimization in Deep Reinforcement Learning27 Apr 2023 0 repositories listed
-
BCQQ: Batch-Constraint Quantum Q-Learning with Cyclic Data Re-uploading27 Apr 2023 0 repositories listed
-
Exploring the flavor structure of quarks and leptons with reinforcement learning27 Apr 2023 0 repositories listed
-
One-Step Distributional Reinforcement Learning27 Apr 2023 0 repositories listed
-
Learning Agile Soccer Skills for a Bipedal Robot with Deep Reinforcement Learning26 Apr 2023 0 repositories listed
-
Multi-criteria Hardware Trojan Detection: A Reinforcement Learning Approach26 Apr 2023 0 repositories listed
-
Reinforcement Learning with Partial Parametric Model Knowledge26 Apr 2023 0 repositories listed
-
A optimization framework for herbal prescription planning based on deep reinforcement learning25 Apr 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.