Browse State-of-the-Art › Reinforcement Learning › Papers, page 46
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 46 of 132: papers 4,501 to 4,600 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Focusing Robot Open-Ended Reinforcement Learning Through Users' Purposes16 Mar 2025 0 repositories listed
-
Evaluation-Time Policy Switching for Offline Reinforcement Learning15 Mar 2025 0 repositories listed
-
Hierarchical Reinforcement Learning for Safe Mapless Navigation with Congestion Estimation15 Mar 2025 0 repositories listed
-
A Review of DeepSeek Models' Key Innovative Techniques14 Mar 2025 0 repositories listed
-
Contextual Similarity Distillation: Ensemble Uncertainties with a Single Model14 Mar 2025 0 repositories listed
-
Dynamic Obstacle Avoidance with Bounded Rationality Adversarial Reinforcement Learning14 Mar 2025 0 repositories listed
-
Sketch-to-Skill: Bootstrapping Robot Learning with Human Drawn Trajectory Sketches14 Mar 2025 0 repositories listed
-
ES-Parkour: Advanced Robot Parkour with Bio-inspired Event Camera and Spiking Neural Network13 Mar 2025 0 repositories listed
-
PRISM: Preference Refinement via Implicit Scene Modeling for 3D Vision-Language Preference-Based Reinforcement Learning13 Mar 2025 0 repositories listed
-
Reinforcement Learning and Life Cycle Assessment for a Circular Economy -- Towards Progressive Computer Science13 Mar 2025 0 repositories listed
-
Evaluating Reinforcement Learning Safety and Trustworthiness in Cyber-Physical Systems12 Mar 2025 0 repositories listed
-
MarineGym: A High-Performance Reinforcement Learning Platform for Underwater Robotics12 Mar 2025 0 repositories listed
-
Reinforcement Learning is all You Need12 Mar 2025 0 repositories listed
-
Rule-Guided Reinforcement Learning Policy Evaluation and Improvement12 Mar 2025 0 repositories listed
-
Strategyproof Reinforcement Learning from Human Feedback12 Mar 2025 0 repositories listed
-
Unified Locomotion Transformer with Simultaneous Sim-to-Real Transfer for Quadrupeds12 Mar 2025 0 repositories listed
-
Enhancing Traffic Signal Control through Model-based Reinforcement Learning and Policy Reuse11 Mar 2025 0 repositories listed
-
LangTime: A Language-Guided Unified Model for Time Series Forecasting with Proximal Policy Optimization11 Mar 2025 0 repositories listed
-
Meta-Reinforcement Learning with Discrete World Models for Adaptive Load Balancing11 Mar 2025 0 repositories listed
-
MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models11 Mar 2025 0 repositories listed
-
AuthorMist: Evading AI Text Detectors with Reinforcement Learning10 Mar 2025 0 repositories listed
-
Goal Conditioned Reinforcement Learning for Photo Finishing Tuning10 Mar 2025 0 repositories listed
-
PER-DPP Sampling Framework and Its Application in Path Planning10 Mar 2025 0 repositories listed
-
Reinforcement Learning Based Symbolic Regression for Load Modeling10 Mar 2025 0 repositories listed
-
Research and Design on Intelligent Recognition of Unordered Targets for Robots Based on Reinforcement Learning10 Mar 2025 0 repositories listed
-
Censoring-Aware Tree-Based Reinforcement Learning for Estimating Dynamic Treatment Regimes with Censored Outcomes9 Mar 2025 0 repositories listed
-
Precise Insulin Delivery for Artificial Pancreas: A Reinforcement Learning Optimized Adaptive Fuzzy Control Approach9 Mar 2025 0 repositories listed
-
Probabilistic Shielding for Safe Reinforcement Learning9 Mar 2025 0 repositories listed
-
Impoola: The Power of Average Pooling for Image-Based Deep Reinforcement Learning7 Mar 2025 0 repositories listed
-
Multi-Robot Collaboration through Reinforcement Learning and Abstract Simulation7 Mar 2025 0 repositories listed
-
Multi-Task Reinforcement Learning Enables Parameter Scaling7 Mar 2025 0 repositories listed
-
Energy-Weighted Flow Matching for Offline Reinforcement Learning6 Mar 2025 0 repositories listed
-
Hedging with Sparse Reward Reinforcement Learning6 Mar 2025 0 repositories listed
-
Knowledge Retention for Continual Model-Based Reinforcement Learning6 Mar 2025 0 repositories listed
-
Multi-Agent Inverse Q-Learning from Demonstrations6 Mar 2025 0 repositories listed
-
Provably Correct Automata Embeddings for Optimal Automata-Conditioned Reinforcement Learning6 Mar 2025 0 repositories listed
-
Towards Autonomous Reinforcement Learning for Real-World Robotic Manipulation with Large Language Models6 Mar 2025 0 repositories listed
-
Human Implicit Preference-Based Policy Fine-tuning for Multi-Agent Reinforcement Learning in USV Swarm5 Mar 2025 0 repositories listed
-
Probabilistic Insights for Efficient Exploration Strategies in Reinforcement Learning5 Mar 2025 0 repositories listed
-
Closing the Intent-to-Behavior Gap via Fulfillment Priority Logic4 Mar 2025 0 repositories listed
-
Reinforcement Learning-based Threat Assessment4 Mar 2025 0 repositories listed
-
Multi-Agent Reinforcement Learning with Long-Term Performance Objectives for Service Workforce Optimization3 Mar 2025 0 repositories listed
-
Adversarial Agents: Black-Box Evasion Attacks with Reinforcement Learning3 Mar 2025 0 repositories listed
-
Active Alignments of Lens Systems with Reinforcement Learning3 Mar 2025 0 repositories listed
-
CE-U: Cross Entropy Unlearning3 Mar 2025 0 repositories listed
-
Differentiable Information Enhanced Model-Based Reinforcement Learning3 Mar 2025 0 repositories listed
-
DPR: Diffusion Preference-based Reward for Offline Reinforcement Learning3 Mar 2025 0 repositories listed
-
Improving Plasticity in Non-stationary Reinforcement Learning with Evidential Proximal Policy Optimization3 Mar 2025 0 repositories listed
-
Stone Soup Multi-Target Tracking Feature Extraction For Autonomous Search And Track In Deep Reinforcement Learning Environment3 Mar 2025 0 repositories listed
-
M3HF: Multi-agent Reinforcement Learning from Multi-phase Human Feedback of Mixed Quality3 Mar 2025 0 repositories listed
-
The Emergence of Grammar through Reinforcement Learning3 Mar 2025 0 repositories listed
-
LADDER: Self-Improving LLMs Through Recursive Problem Decomposition2 Mar 2025 0 repositories listed
-
Minimax Optimal Reinforcement Learning with Quasi-Optimism2 Mar 2025 0 repositories listed
-
Adaptive Entanglement Routing with Deep Q-Networks in Quantum Networks1 Mar 2025 0 repositories listed
-
Scalable Reinforcement Learning for Virtual Machine Scheduling1 Mar 2025 0 repositories listed
-
Shaping Laser Pulses with Reinforcement Learning1 Mar 2025 0 repositories listed
-
Robust Gymnasium: A Unified Modular Benchmark for Robust Reinforcement Learning27 Feb 2025 0 repositories listed
-
Sim-to-Real Reinforcement Learning for Vision-Based Dexterous Manipulation on Humanoids27 Feb 2025 0 repositories listed
-
Combining Planning and Reinforcement Learning for Solving Relational Multiagent Domains26 Feb 2025 0 repositories listed
-
Efficient Reinforcement Learning by Guiding Generalist World Models with Non-Curated Data26 Feb 2025 0 repositories listed
-
Recurrent Auto-Encoders for Enhanced Deep Reinforcement Learning in Wilderness Search and Rescue Planning26 Feb 2025 0 repositories listed
-
Adaptive Nesterov Accelerated Distributional Deep Hedging for Efficient Volatility Risk Management25 Feb 2025 0 repositories listed
-
Applications of deep reinforcement learning to urban transit network design25 Feb 2025 0 repositories listed
-
CayleyPy RL: Pathfinding and Reinforcement Learning on Cayley Graphs25 Feb 2025 0 repositories listed
-
MPO: An Efficient Post-Processing Framework for Mixing Diverse Preference Alignment25 Feb 2025 0 repositories listed
-
Survey on Strategic Mining in Blockchain: A Reinforcement Learning Approach24 Feb 2025 0 repositories listed
-
Towards Reinforcement Learning for Exploration of Speculative Execution Vulnerabilities24 Feb 2025 0 repositories listed
-
Exploring Sentiment Manipulation by LLM-Enabled Intelligent Trading Agents22 Feb 2025 0 repositories listed
-
Towards User-level Private Reinforcement Learning with Human Feedback22 Feb 2025 0 repositories listed
-
The Evolving Landscape of LLM- and VLM-Integrated Reinforcement Learning21 Feb 2025 0 repositories listed
-
Towards a Reward-Free Reinforcement Learning Framework for Vehicle Control21 Feb 2025 0 repositories listed
-
Causal Mean Field Multi-Agent Reinforcement Learning20 Feb 2025 0 repositories listed
-
Is Q-learning an Ill-posed Problem?20 Feb 2025 0 repositories listed
-
μRL: Discovering Transient Execution Vulnerabilities Using Reinforcement Learning20 Feb 2025 0 repositories listed
-
SPRIG: Stackelberg Perception-Reinforcement Learning with Internal Game Dynamics20 Feb 2025 0 repositories listed
-
Multi-Target Radar Search and Track Using Sequence-Capable Deep Reinforcement Learning19 Feb 2025 0 repositories listed
-
Uncertainty quantification for Markov chains with application to temporal difference learning19 Feb 2025 0 repositories listed
-
Continuous Learning Conversational AI: A Personalized Agent Framework via A2C Reinforcement Learning18 Feb 2025 0 repositories listed
-
Implicit Repair with Reinforcement Learning in Emergent Communication18 Feb 2025 0 repositories listed
-
Self-Supervised Transformers as Iterative Solution Improvers for Constraint Satisfaction18 Feb 2025 0 repositories listed
-
Theorem Prover as a Judge for Synthetic Data Generation18 Feb 2025 0 repositories listed
-
FitLight: Federated Imitation Learning for Plug-and-Play Autonomous Traffic Signal Control17 Feb 2025 0 repositories listed
-
Intelligent Mobile AI-Generated Content Services via Interactive Prompt Engineering and Dynamic Service Provisioning17 Feb 2025 0 repositories listed
-
Learning to Reason at the Frontier of Learnability17 Feb 2025 0 repositories listed
-
Theoretical Barriers in Bellman-Based Reinforcement Learning17 Feb 2025 0 repositories listed
-
Solving Online Resource-Constrained Scheduling for Follow-Up Observation in Astronomy: a Reinforcement Learning Approach16 Feb 2025 0 repositories listed
-
A Tutorial on LLM Reasoning: Relevant Methods behind ChatGPT o115 Feb 2025 0 repositories listed
-
Rule-Bottleneck Reinforcement Learning: Joint Explanation and Decision Optimization for Resource Allocation with Language Agents15 Feb 2025 0 repositories listed
-
Tackling the Zero-Shot Reinforcement Learning Loss Directly15 Feb 2025 0 repositories listed
-
Causal Information Prioritization for Efficient Reinforcement Learning14 Feb 2025 0 repositories listed
-
Combinatorial Reinforcement Learning with Preference Feedback14 Feb 2025 0 repositories listed
-
Do We Need to Verify Step by Step? Rethinking Process Supervision from a Theoretical Perspective14 Feb 2025 0 repositories listed
-
Dynamic Reinforcement Learning for Actors14 Feb 2025 0 repositories listed
-
Reinforcement Learning based Constrained Optimal Control: an Interpretable Reward Design14 Feb 2025 0 repositories listed
-
Reinforcement Learning in Strategy-Based and Atari Games: A Review of Google DeepMinds Innovations14 Feb 2025 0 repositories listed
-
Self-Consistent Model-based Adaptation for Visual Reinforcement Learning14 Feb 2025 0 repositories listed
-
A Survey of Reinforcement Learning for Optimization in Automation13 Feb 2025 0 repositories listed
-
Analysis of Off-Policy n-Step TD-Learning with Linear Function Approximation13 Feb 2025 0 repositories listed
-
Coupled Rendezvous and Docking Maneuver control of satellite using Reinforcement learning-based Adaptive Fixed-Time Sliding Mode Controller13 Feb 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.