Browse State-of-the-Art › Reinforcement Learning › Papers, page 47
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 47 of 132: papers 4,601 to 4,700 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Variable Stiffness for Robust Locomotion through Reinforcement Learning13 Feb 2025 0 repositories listed
-
Deep Reinforcement Learning-Based User Scheduling for Collaborative Perception12 Feb 2025 0 repositories listed
-
Provably Robust Federated Reinforcement Learning12 Feb 2025 0 repositories listed
-
A Survey of In-Context Reinforcement Learning11 Feb 2025 0 repositories listed
-
Exploratory Diffusion Model for Unsupervised Reinforcement Learning11 Feb 2025 0 repositories listed
-
Logarithmic Regret for Online KL-Regularized Reinforcement Learning11 Feb 2025 0 repositories listed
-
Optimal Actuator Attacks on Autonomous Vehicles Using Reinforcement Learning11 Feb 2025 0 repositories listed
-
PICTS: A Novel Deep Reinforcement Learning Approach for Dynamic P-I Control in Scanning Probe Microscopy11 Feb 2025 0 repositories listed
-
Polynomial-Time Approximability of Constrained Reinforcement Learning11 Feb 2025 0 repositories listed
-
Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization11 Feb 2025 0 repositories listed
-
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning11 Feb 2025 0 repositories listed
-
Intelligent Offloading in Vehicular Edge Computing: A Comprehensive Review of Deep Reinforcement Learning Approaches and Architectures10 Feb 2025 0 repositories listed
-
A Survey on Explainable Deep Reinforcement Learning8 Feb 2025 0 repositories listed
-
Real Time Control of Tandem-Wing Experimental Platform Using Concerto Reinforcement Learning8 Feb 2025 0 repositories listed
-
Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning8 Feb 2025 0 repositories listed
-
Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning7 Feb 2025 0 repositories listed
-
Agency Is Frame-Dependent6 Feb 2025 0 repositories listed
-
Behavioral Entropy-Guided Dataset Generation for Offline Reinforcement Learning6 Feb 2025 0 repositories listed
-
Fairness Aware Reinforcement Learning via Proximal Policy Optimization6 Feb 2025 0 repositories listed
-
Reinforcement Learning on Dyads to Enhance Medication Adherence6 Feb 2025 0 repositories listed
-
Double Distillation Network for Multi-Agent Reinforcement Learning5 Feb 2025 0 repositories listed
-
OmniRL: In-Context Reinforcement Learning by Large-Scale Meta-Training in Randomized Worlds5 Feb 2025 0 repositories listed
-
Teaching Language Models to Critique via Reinforcement Learning5 Feb 2025 0 repositories listed
-
CH-MARL: Constrained Hierarchical Multiagent Reinforcement Learning for Sustainable Maritime Logistics4 Feb 2025 0 repositories listed
-
DHP: Discrete Hierarchical Planning for Hierarchical Reinforcement Learning Agents4 Feb 2025 0 repositories listed
-
DIME:Diffusion-Based Maximum Entropy Reinforcement Learning4 Feb 2025 0 repositories listed
-
Policy-Guided Causal State Representation for Offline Reinforcement Learning Recommendation4 Feb 2025 0 repositories listed
-
ACECODER: Acing Coder RL via Automated Test-Case Synthesis3 Feb 2025 0 repositories listed
-
Competitive Programming with Large Reasoning Models3 Feb 2025 0 repositories listed
-
Exploratory Utility Maximization Problem with Tsallis Entropy3 Feb 2025 0 repositories listed
-
Generalized Lanczos method for systematic optimization of neural-network quantum states3 Feb 2025 0 repositories listed
-
Preference VLM: Leveraging VLMs for Scalable Preference-Based Reinforcement Learning3 Feb 2025 0 repositories listed
-
Process-Supervised Reinforcement Learning for Code Generation3 Feb 2025 0 repositories listed
-
Reinforcement Learning for Long-Horizon Interactive LLM Agents3 Feb 2025 0 repositories listed
-
Reinforcement Learning with Segment Feedback3 Feb 2025 0 repositories listed
-
The Differences Between Direct Alignment Algorithms are a Blur3 Feb 2025 0 repositories listed
-
Toward Task Generalization via Memory Augmentation in Meta-Reinforcement Learning3 Feb 2025 0 repositories listed
-
VR-Robo: A Real-to-Sim-to-Real Framework for Visual Robot Navigation and Locomotion3 Feb 2025 0 repositories listed
-
Compositional Concept-Based Neuron-Level Interpretability for Deep Reinforcement Learning2 Feb 2025 0 repositories listed
-
Decorrelated Soft Actor-Critic for Efficient Deep Reinforcement Learning31 Jan 2025 0 repositories listed
-
Shaping Sparse Rewards in Reinforcement Learning: A Semi-supervised Approach31 Jan 2025 0 repositories listed
-
SpikingSoft: A Spiking Neuron Controller for Bio-inspired Locomotion with Soft Snake Robots31 Jan 2025 0 repositories listed
-
Hybrid Group Relative Policy Optimization: A Multi-Sample Approach to Enhancing Policy Optimization30 Jan 2025 0 repositories listed
-
Certificated Actor-Critic: Hierarchical Reinforcement Learning with Control Barrier Functions for Safe Navigation29 Jan 2025 0 repositories listed
-
Digital Twin Synchronization: Bridging the Sim-RL Agent to a Real-Time Robotic Additive Manufacturing Control29 Jan 2025 0 repositories listed
-
Reinforcement-Learning Portfolio Allocation with Dynamic Embedding of Market Information29 Jan 2025 0 repositories listed
-
The M-factor: A Novel Metric for Evaluating Neural Architecture Search in Resource-Constrained Environments29 Jan 2025 0 repositories listed
-
Applying Ensemble Models based on Graph Neural Network and Reinforcement Learning for Wind Power Forecasting28 Jan 2025 0 repositories listed
-
Improving Vision-Language-Action Model with Online Reinforcement Learning28 Jan 2025 0 repositories listed
-
Induced Modularity and Community Detection for Functionally Interpretable Reinforcement Learning28 Jan 2025 0 repositories listed
-
On the Interplay Between Sparsity and Training in Deep Reinforcement Learning28 Jan 2025 0 repositories listed
-
Safe Reinforcement Learning for Real-World Engine Control28 Jan 2025 0 repositories listed
-
Generative AI for Lyapunov Optimization Theory in UAV-based Low-Altitude Economy Networking27 Jan 2025 0 repositories listed
-
Reinforcement Learning for Quantum Circuit Design: Using Matrix Representations27 Jan 2025 0 repositories listed
-
Selective Experience Sharing in Reinforcement Learning Enhances Interference Management27 Jan 2025 0 repositories listed
-
27 Jan 2025 0 repositories listed Syntology 5 ran (of which 4 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 4 unverified (of 9 harvested samples) · 4 pointer-only (licence)
-
Advancing TDFN: Precise Fixation Point Generation Using Reconstruction Differences26 Jan 2025 0 repositories listed
-
Contextual Knowledge Sharing in Multi-Agent Reinforcement Learning with Decentralized Communication and Coordination26 Jan 2025 0 repositories listed
-
Data Center Cooling System Optimization Using Offline Reinforcement Learning25 Jan 2025 0 repositories listed
-
Extensive Exploration in Complex Traffic Scenarios using Hierarchical Reinforcement Learning25 Jan 2025 0 repositories listed
-
Faster Machine Translation Ensembling with Reinforcement Learning and Competitive Correction25 Jan 2025 0 repositories listed
-
Music Generation using Human-In-The-Loop Reinforcement Learning25 Jan 2025 0 repositories listed
-
Predictive Lagrangian Optimization for Constrained Reinforcement Learning25 Jan 2025 0 repositories listed
-
Reinforcement Learning for Efficient Returns Management24 Jan 2025 0 repositories listed
-
RL + Transformer = A General-Purpose Problem Solver24 Jan 2025 0 repositories listed
-
Towards Efficient Multi-Objective Optimisation for Real-World Power Grid Topology Control24 Jan 2025 0 repositories listed
-
Reinforcement Learning Platform for Adversarial Black-box Attacks with Custom Distortion Filters23 Jan 2025 0 repositories listed
-
Attention-Driven Hierarchical Reinforcement Learning with Particle Filtering for Source Localization in Dynamic Fields22 Jan 2025 0 repositories listed
-
Deep Reinforcement Learning with Hybrid Intrinsic Reward Model22 Jan 2025 0 repositories listed
-
Exploring the Technology Landscape through Topic Modeling, Expert Involvement, and Reinforcement Learning22 Jan 2025 0 repositories listed
-
Offline Critic-Guided Diffusion Policy for Multi-User Delay-Constrained Scheduling22 Jan 2025 0 repositories listed
-
RAG-Reward: Optimizing RAG with Reward Modeling and RLHF22 Jan 2025 0 repositories listed
-
UAV-assisted Internet of Vehicles: A Framework Empowered by Reinforcement Learning and Blockchain22 Jan 2025 0 repositories listed
-
ARM-IRL: Adaptive Resilience Metric Quantification Using Inverse Reinforcement Learning21 Jan 2025 0 repositories listed
-
Compositional Instruction Following with Language Models and Reinforcement Learning21 Jan 2025 0 repositories listed
-
DNRSelect: Active Best View Selection for Deferred Neural Rendering21 Jan 2025 0 repositories listed
-
Group-Agent Reinforcement Learning with Heterogeneous Agents21 Jan 2025 0 repositories listed
-
Heuristic Deep Reinforcement Learning for Phase Shift Optimization in RIS-assisted Secure Satellite Communication Systems with RSMA21 Jan 2025 0 repositories listed
-
Reinforcement Learning Constrained Beam Search for Parameter Optimization of Paper Drying Under Flexible Constraints21 Jan 2025 0 repositories listed
-
Secure Resource Allocation via Constrained Deep Reinforcement Learning20 Jan 2025 0 repositories listed
-
The impact of intrinsic rewards on exploration in Reinforcement Learning20 Jan 2025 0 repositories listed
-
Blockchain-assisted Demonstration Cloning for Multi-Agent Deep Reinforcement Learning19 Jan 2025 0 repositories listed
-
FORLAPS: An Innovative Data-Driven Reinforcement Learning Approach for Prescriptive Process Monitoring17 Jan 2025 0 repositories listed
-
Wasserstein Adaptive Value Estimation for Actor-Critic Reinforcement Learning17 Jan 2025 0 repositories listed
-
EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning16 Jan 2025 0 repositories listed
-
Fast Searching of Extreme Operating Conditions for Relay Protection Setting Calculation Based on Graph Neural Network and Reinforcement Learning16 Jan 2025 0 repositories listed
-
A Reinforcement Learning Approach to Quiet and Safe UAM Traffic Management15 Jan 2025 0 repositories listed
-
Application of Deep Reinforcement Learning to UAV Swarming for Ground Surveillance15 Jan 2025 0 repositories listed
-
Average-Reward Reinforcement Learning with Entropy Regularization15 Jan 2025 0 repositories listed
-
Inferring Transition Dynamics from Value Functions15 Jan 2025 0 repositories listed
-
Reinforcement Learning-Enhanced Procedural Generation for Dynamic Narrative-Driven AR Experiences15 Jan 2025 0 repositories listed
-
Dynamic Pricing in High-Speed Railways Using Multi-Agent Reinforcement Learning14 Jan 2025 0 repositories listed
-
Optimization of Link Configuration for Satellite Communication Using Reinforcement Learning14 Jan 2025 0 repositories listed
-
READ: Reinforcement-based Adversarial Learning for Text Classification with Limited Labeled Data14 Jan 2025 0 repositories listed
-
Online inductive learning from answer sets for efficient reinforcement learning exploration13 Jan 2025 0 repositories listed
-
Performance Optimization of Ratings-Based Reinforcement Learning13 Jan 2025 0 repositories listed
-
TIMRL: A Novel Meta-Reinforcement Learning Framework for Non-Stationary and Multi-Task Environments13 Jan 2025 0 repositories listed
-
Average Reward Reinforcement Learning for Wireless Radio Resource Management12 Jan 2025 0 repositories listed
-
Pareto Set Learning for Multi-Objective Reinforcement Learning12 Jan 2025 0 repositories listed
-
Hierarchical Reinforcement Learning for Optimal Agent Grouping in Cooperative Systems11 Jan 2025 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.