Browse State-of-the-Art › Reinforcement Learning › Papers, page 60
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 60 of 132: papers 5,901 to 6,000 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
A Central Motor System Inspired Pre-training Reinforcement Learning for Robotic Control14 Nov 2023 0 repositories listed
-
Adversarial Imitation Learning On Aggregated Data14 Nov 2023 0 repositories listed
-
On-Policy Policy Gradient Reinforcement Learning Without On-Policy Sampling14 Nov 2023 0 repositories listed
-
When Mining Electric Locomotives Meet Reinforcement Learning14 Nov 2023 0 repositories listed
-
A Large Deviations Perspective on Policy Gradient Algorithms13 Nov 2023 0 repositories listed
-
An introduction to reinforcement learning for neuroscience13 Nov 2023 0 repositories listed
-
TIAGo RL: Simulated Reinforcement Learning Environments with Tactile Data for Mobile Robots13 Nov 2023 0 repositories listed
-
Towards Continual Reinforcement Learning for Quadruped Robots12 Nov 2023 0 repositories listed
-
Causal Inference on Investment Constraints and Non-stationarity in Dynamic Portfolio Optimization through Reinforcement Learning8 Nov 2023 0 repositories listed
-
Real-Time Recurrent Reinforcement Learning8 Nov 2023 0 repositories listed
-
Reinforcement Learning Generalization for Nonlinear Systems Through Dual-Scale Homogeneity Transformations8 Nov 2023 0 repositories listed
-
A Method to Improve the Performance of Reinforcement Learning Based on the Y Operator for a Class of Stochastic Differential Equation-Based Child-Mother Systems7 Nov 2023 0 repositories listed
-
A Novel Variational Lower Bound for Inverse Reinforcement Learning7 Nov 2023 0 repositories listed
-
Hypothesis Network Planned Exploration for Rapid Meta-Reinforcement Learning Adaptation7 Nov 2023 0 repositories listed
-
Reinforcement Twinning: from digital twins to model-based reinforcement learning7 Nov 2023 0 repositories listed
-
Stable Modular Control via Contraction Theory for Reinforcement Learning7 Nov 2023 0 repositories listed
-
Time-Efficient Reinforcement Learning with Stochastic Stateful Policies7 Nov 2023 0 repositories listed
-
Environmental-Impact Based Multi-Agent Reinforcement Learning6 Nov 2023 0 repositories listed
-
Kindness in Multi-Agent Reinforcement Learning6 Nov 2023 0 repositories listed
-
CIRCLE: Multi-Turn Query Clarifications with Reinforcement Learning5 Nov 2023 0 repositories listed
-
Staged Reinforcement Learning for Complex Tasks through Decomposed Environments5 Nov 2023 0 repositories listed
-
Accelerating Reinforcement Learning of Robotic Manipulations via Feedback from Large Language Models4 Nov 2023 0 repositories listed
-
DeliverAI: Reinforcement Learning Based Distributed Path-Sharing Network for Food Deliveries3 Nov 2023 0 repositories listed
-
Epidemic Decision-making System Based Federated Reinforcement Learning3 Nov 2023 0 repositories listed
-
Imitation Bootstrapped Reinforcement Learning3 Nov 2023 0 repositories listed
-
LLMs-augmented Contextual Bandit3 Nov 2023 0 repositories listed
-
Robust Adversarial Reinforcement Learning via Bounded Rationality Curricula3 Nov 2023 0 repositories listed
-
SMORE: Score Models for Offline Goal-Conditioned Reinforcement Learning3 Nov 2023 0 repositories listed
-
Toward Reinforcement Learning-based Rectilinear Macro Placement Under Human Constraints3 Nov 2023 0 repositories listed
-
Anytime-Competitive Reinforcement Learning with Policy Prior2 Nov 2023 0 repositories listed
-
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing2 Nov 2023 0 repositories listed
-
Dynamic Fair Federated Learning Based on Reinforcement Learning2 Nov 2023 0 repositories listed
-
Rethinking Decision Transformer via Hierarchical Reinforcement Learning1 Nov 2023 0 repositories listed
-
SCPO: Safe Reinforcement Learning with Safety Critic Policy Optimization1 Nov 2023 0 repositories listed
-
Autonomous Robotic Reinforcement Learning with Asynchronous Human Feedback31 Oct 2023 0 repositories listed
-
Dropout Strategy in Reinforcement Learning: Limiting the Surrogate Objective Variance in Policy Optimization Methods31 Oct 2023 0 repositories listed
-
Safe multi-agent motion planning under uncertainty for drones using filtered reinforcement learning31 Oct 2023 0 repositories listed
-
Towards Instance-Optimality in Online PAC Reinforcement Learning31 Oct 2023 0 repositories listed
-
Adversarial Batch Inverse Reinforcement Learning: Learn to Reward from Imperfect Demonstration for Interactive Recommendation30 Oct 2023 0 repositories listed
-
Efficient Exploration in Continuous-time Model-based Reinforcement Learning30 Oct 2023 0 repositories listed
-
Improved Bayesian Regret Bounds for Thompson Sampling in Reinforcement Learning30 Oct 2023 0 repositories listed
-
Remember what you did so you know what to do next30 Oct 2023 0 repositories listed
-
Automaton Distillation: Neuro-Symbolic Transfer Learning for Deep Reinforcement Learning29 Oct 2023 0 repositories listed
-
MAG-GNN: Reinforcement Learning Boosted Graph Neural Network29 Oct 2023 0 repositories listed
-
Gen2Sim: Scaling up Robot Learning in Simulation with Generative Models27 Oct 2023 0 repositories listed
-
Guided Data Augmentation for Offline Reinforcement Learning and Imitation Learning27 Oct 2023 0 repositories listed
-
CQM: Curriculum Reinforcement Learning with a Quantized World Model26 Oct 2023 0 repositories listed
-
Demonstration-Regularized RL26 Oct 2023 0 repositories listed
-
Relational Object-Centric Actor-Critic26 Oct 2023 0 repositories listed
-
Large Language Models as Generalizable Policies for Embodied Tasks26 Oct 2023 0 repositories listed
-
26 Oct 2023 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 0 honoured, 0 violated, 3 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 4 harvested samples) · 1 pointer-only (licence)
-
AI Agent as Urban Planner: Steering Stakeholder Dynamics in Urban Planning via Consensus-based Multi-Agent Reinforcement Learning25 Oct 2023 0 repositories listed
-
Controlled Decoding from Language Models25 Oct 2023 0 repositories listed
-
Imperfect Digital Twin Assisted Low Cost Reinforcement Training for Multi-UAV Networks25 Oct 2023 0 repositories listed
-
Model-enhanced Contrastive Reinforcement Learning for Sequential Recommendation25 Oct 2023 0 repositories listed
-
MultiPrompter: Cooperative Prompt Optimization with Multi-Agent Reinforcement Learning25 Oct 2023 0 repositories listed
-
Pitfall of Optimism: Distributional Reinforcement Learning by Randomizing Risk Criterion25 Oct 2023 0 repositories listed
-
Privately Aligning Language Models with Reinforcement Learning25 Oct 2023 0 repositories listed
-
Reinforcement Learning for SBM Graphon Games with Re-Sampling25 Oct 2023 0 repositories listed
-
Policy Optimization via Adv2: Adversarial Learning on Advantage Functions25 Oct 2023 0 repositories listed
-
Towards Control-Centric Representations in Reinforcement Learning from Images25 Oct 2023 0 repositories listed
-
A Contextualized Real-Time Multimodal Emotion Recognition for Conversational Agents using Graph Convolutional Networks in Reinforcement Learning24 Oct 2023 0 repositories listed
-
COPR: Continual Learning Human Preference through Optimal Policy Regularization24 Oct 2023 0 repositories listed
-
Reinforcement learning based local path planning for mobile robot24 Oct 2023 0 repositories listed
-
A Doubly Robust Approach to Sparse Reinforcement Learning23 Oct 2023 0 repositories listed
-
Active teacher selection for reinforcement learning from human feedback23 Oct 2023 0 repositories listed
-
AI on the Water: Applying DRL to Autonomous Vessel Navigation23 Oct 2023 0 repositories listed
-
Comparison of path following in ships using modern and traditional controllers23 Oct 2023 0 repositories listed
-
Diverse Priors for Deep Reinforcement Learning23 Oct 2023 0 repositories listed
-
Enhancing Robotic Manipulation: Harnessing the Power of Multi-Task Reinforcement Learning and Single Life Reinforcement Learning in Meta-World23 Oct 2023 0 repositories listed
-
Robot Fine-Tuning Made Easy: Pre-Training Rewards and Policies for Autonomous Real-World Reinforcement Learning23 Oct 2023 0 repositories listed
-
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL23 Oct 2023 0 repositories listed
-
One is More: Diverse Perspectives within a Single Network for Efficient DRL21 Oct 2023 0 repositories listed
-
Interpretable Deep Reinforcement Learning for Optimizing Heterogeneous Energy Storage Systems20 Oct 2023 0 repositories listed
-
Reward Shaping for Happier Autonomous Cyber Security Agents20 Oct 2023 0 repositories listed
-
Tree Search in DAG Space with Model-based Reinforcement Learning for Causal Discovery20 Oct 2023 0 repositories listed
-
Safety-Gymnasium: A Unified Safe Reinforcement Learning Benchmark19 Oct 2023 0 repositories listed
-
Accelerate Presolve in Large-Scale Linear Programming via Reinforcement Learning18 Oct 2023 0 repositories listed
-
Action-Quantized Offline Reinforcement Learning for Robotic Skill Learning18 Oct 2023 0 repositories listed
-
Fact-based Agent modeling for Multi-Agent Reinforcement Learning18 Oct 2023 0 repositories listed
-
Improving Generalization of Alignment with Human Preferences through Group Invariant Learning18 Oct 2023 0 repositories listed
-
On The Expressivity of Objective-Specification Formalisms in Reinforcement Learning18 Oct 2023 0 repositories listed
-
Quantum Speedups in Regret Analysis of Infinite Horizon Average-Reward Markov Decision Processes18 Oct 2023 0 repositories listed
-
Combat Urban Congestion via Collaboration: Heterogeneous GNN-based MARL for Coordinated Platooning and Traffic Signal Control17 Oct 2023 0 repositories listed
-
Neural Packing: from Visual Sensing to Reinforcement Learning17 Oct 2023 0 repositories listed
-
End-to-end Offline Reinforcement Learning for Glycemia Control16 Oct 2023 0 repositories listed
-
Leveraging Topological Maps in Deep Reinforcement Learning for Multi-Object Navigation16 Oct 2023 0 repositories listed
-
Mimicking the Maestro: Exploring the Efficacy of a Virtual AI Teacher in Fine Motor Skill Acquisition16 Oct 2023 0 repositories listed
-
Verbosity Bias in Preference Labeling by Large Language Models16 Oct 2023 0 repositories listed
-
Deep Reinforcement Learning with Explicit Context Representation15 Oct 2023 0 repositories listed
-
Federated Reinforcement Learning for Resource Allocation in V2X Networks15 Oct 2023 0 repositories listed
-
Robust Multi-Agent Reinforcement Learning by Mutual Information Regularization15 Oct 2023 0 repositories listed
-
A Blockchain-empowered Multi-Aggregator Federated Learning Architecture in Edge Computing with Deep Reinforcement Learning Optimization14 Oct 2023 0 repositories listed
-
Automatic Music Playlist Generation via Simulation-based Reinforcement Learning13 Oct 2023 0 repositories listed
-
Goodhart's Law in Reinforcement Learning13 Oct 2023 0 repositories listed
-
Offline Reinforcement Learning for Optimizing Production Bidding Policies13 Oct 2023 0 repositories listed
-
Novelty Detection in Reinforcement Learning with World Models12 Oct 2023 0 repositories listed
-
Aligning Data Selection with Performance: Performance-driven Reinforcement Learning for Active Learning in Object Detection12 Oct 2023 0 repositories listed
-
Reinforcement Learning of Display Transfer Robots in Glass Flow Control Systems: A Physical Simulation-Based Approach12 Oct 2023 0 repositories listed
-
Deep Reinforcement Learning for Autonomous Cyber Defence: A Survey11 Oct 2023 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.