Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 63
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 63 of 152: papers 6,201 to 6,300 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning3 May 2024 0 repositories listed
-
Learning Robust Autonomous Navigation and Locomotion for Wheeled-Legged Robots3 May 2024 0 repositories listed
-
Model-based reinforcement learning for protein backbone design3 May 2024 0 repositories listed
-
Natural Policy Gradient and Actor Critic Methods for Constrained Multi-Task Reinforcement Learning3 May 2024 0 repositories listed
-
Zero-Sum Positional Differential Games as a Framework for Robust Reinforcement Learning: Deep Q-Learning Approach3 May 2024 0 repositories listed
-
Constrained Reinforcement Learning Under Model Mismatch2 May 2024 0 repositories listed
-
FLAME: Factuality-Aware Alignment for Large Language Models2 May 2024 0 repositories listed
-
Learning Force Control for Legged Manipulation2 May 2024 0 repositories listed
-
Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks2 May 2024 0 repositories listed
-
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation2 May 2024 0 repositories listed
-
Reinforcement Learning-Guided Semi-Supervised Learning2 May 2024 0 repositories listed
-
Robust Risk-Sensitive Reinforcement Learning with Conditional Value-at-Risk2 May 2024 0 repositories listed
-
Tabular and Deep Reinforcement Learning for Gittins Index2 May 2024 0 repositories listed
-
Navigating WebAI: Training Agents to Complete Web Tasks with Large Language Models and Reinforcement Learning1 May 2024 0 repositories listed
-
Queue-based Eco-Driving at Roundabouts with Reinforcement Learning1 May 2024 0 repositories listed
-
Leveraging Sub-Optimal Data for Human-in-the-Loop Reinforcement Learning30 Apr 2024 0 repositories listed
-
Towards Generalist Robot Learning from Internet Video: A Survey30 Apr 2024 0 repositories listed
-
Control Policy Correction Framework for Reinforcement Learning-based Energy Arbitrage Strategies29 Apr 2024 0 repositories listed
-
Reinforcement Learning Problem Solving with Large Language Models29 Apr 2024 0 repositories listed
-
Sample-Efficient Robust Multi-Agent Reinforcement Learning in the Face of Environmental Uncertainty29 Apr 2024 0 repositories listed
-
Towards Generalizable Agents in Text-Based Educational Environments: A Study of Integrating RL with LLMs29 Apr 2024 0 repositories listed
-
EEG_RL-Net: Enhancing EEG MI Classification through Reinforcement Learning-Optimised Graph Neural Networks26 Apr 2024 0 repositories listed
-
Enhancing Privacy and Security of Autonomous UAV Navigation26 Apr 2024 0 repositories listed
-
Generalize by Touching: Tactile Ensemble Skill Transfer for Robotic Furniture Assembly26 Apr 2024 0 repositories listed
-
Knowledge Transfer for Cross-Domain Reinforcement Learning: A Systematic Review26 Apr 2024 0 repositories listed
-
Offline Reinforcement Learning with Behavioral Supervisor Tuning25 Apr 2024 0 repositories listed
-
Structured Reinforcement Learning for Delay-Optimal Data Transmission in Dense mmWave Networks25 Apr 2024 0 repositories listed
-
ActiveRIR: Active Audio-Visual Exploration for Acoustic Environment Modeling24 Apr 2024 0 repositories listed
-
24 Apr 2024 0 repositories listed Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
GRSN: Gated Recurrent Spiking Neurons for POMDPs and MARL24 Apr 2024 0 repositories listed
-
An MRP Formulation for Supervised Learning: Generalized Temporal Difference Learning Models23 Apr 2024 0 repositories listed
-
Impedance Matching: Enabling an RL-Based Running Jump in a Quadruped Robot23 Apr 2024 0 repositories listed
-
Using deep reinforcement learning to promote sustainable human behaviour on a common pool resource problem23 Apr 2024 0 repositories listed
-
Beyond the Edge: An Advanced Exploration of Reinforcement Learning for Mobile Edge Computing, its Applications, and Future Research Trajectories22 Apr 2024 0 repositories listed
-
Explicit Lipschitz Value Estimation Enhances Policy Robustness Against Perturbation22 Apr 2024 0 repositories listed
-
Fairness Incentives in Response to Unfair Dynamic Pricing22 Apr 2024 0 repositories listed
-
An Offline Reinforcement Learning Algorithm Customized for Multi-Task Fusion in Large-Scale Recommender Systems19 Apr 2024 0 repositories listed
-
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty19 Apr 2024 0 repositories listed
-
Reinforcement Learning Approach for Integrating Compressed Contexts into Knowledge Graphs19 Apr 2024 0 repositories listed
-
Data-Incremental Continual Offline Reinforcement Learning19 Apr 2024 0 repositories listed
-
Actor-Critic Reinforcement Learning with Phased Actor18 Apr 2024 0 repositories listed
-
LTL-Constrained Policy Optimization with Cycle Experience Replay17 Apr 2024 0 repositories listed
-
Learn to Tour: Operator Design For Solution Feasibility Mapping in Pickup-and-delivery Traveling Salesman Problem17 Apr 2024 0 repositories listed
-
Physics-informed Actor-Critic for Coordination of Virtual Inertia from Power Distribution Systems17 Apr 2024 0 repositories listed
-
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding17 Apr 2024 0 repositories listed
-
Automated Discovery of Functional Actual Causes in Complex Environments16 Apr 2024 0 repositories listed
-
Offline Trajectory Generalization for Offline Reinforcement Learning16 Apr 2024 0 repositories listed
-
Achieving Constant Regret in Linear Markov Decision Processes16 Apr 2024 0 repositories listed
-
Simplex Decomposition for Portfolio Allocation Constraints in Reinforcement Learning16 Apr 2024 0 repositories listed
-
Autonomous Path Planning for Intercostal Robotic Ultrasound Imaging Using Reinforcement Learning15 Apr 2024 0 repositories listed
-
Effective Reinforcement Learning Based on Structural Information Principles15 Apr 2024 0 repositories listed
-
Higher Replay Ratio Empowers Sample-Efficient Multi-Agent Reinforcement Learning15 Apr 2024 0 repositories listed
-
The Feasibility of Constrained Reinforcement Learning Algorithms: A Tutorial Study15 Apr 2024 0 repositories listed
-
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts14 Apr 2024 0 repositories listed
-
SmartPathfinder: Pushing the Limits of Heuristic Solutions for Vehicle Routing Problem with Drones Using Reinforcement Learning13 Apr 2024 0 repositories listed
-
Efficient Duple Perturbation Robustness in Low-rank MDPs11 Apr 2024 0 repositories listed
-
Enhancing Policy Gradient with the Polyak Step-Size Adaption11 Apr 2024 0 repositories listed
-
FPGA Divide-and-Conquer Placement using Deep Reinforcement Learning11 Apr 2024 0 repositories listed
-
Leveraging Domain-Unlabeled Data in Offline Reinforcement Learning across Two Domains11 Apr 2024 0 repositories listed
-
On the Sample Efficiency of Abstractions and Potential-Based Reward Shaping in Reinforcement Learning11 Apr 2024 0 repositories listed
-
Dual Ensemble Kalman Filter for Stochastic Optimal Control10 Apr 2024 0 repositories listed
-
Reward Learning from Suboptimal Demonstrations with Applications in Surgical Electrocautery10 Apr 2024 0 repositories listed
-
UAV-Assisted Enhanced Coverage and Capacity in Dynamic MU-mMIMO IoT Systems: A Deep Reinforcement Learning Approach10 Apr 2024 0 repositories listed
-
Adaptable Recovery Behaviors in Robotics: A Behavior Trees and Motion Generators(BTMG) Approach for Failure Management9 Apr 2024 0 repositories listed
-
Asynchronous Federated Reinforcement Learning with Policy Gradient Updates: Algorithm Design and Convergence Analysis9 Apr 2024 0 repositories listed
-
Diverse Randomized Value Functions: A Provably Pessimistic Approach for Offline Reinforcement Learning9 Apr 2024 0 repositories listed
-
FGAIF: Aligning Large Vision-Language Models with Fine-grained AI Feedback7 Apr 2024 0 repositories listed
-
Transform then Explore: a Simple and Effective Technique for Exploratory Combinatorial Optimization with Reinforcement Learning6 Apr 2024 0 repositories listed
-
Enhancing IoT Intelligence: A Transformer-based Reinforcement Learning Methodology5 Apr 2024 0 repositories listed
-
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers4 Apr 2024 0 repositories listed
-
Distributionally Robust Reinforcement Learning with Interactive Data Collection: Fundamental Hardness and Near-Optimal Algorithm4 Apr 2024 0 repositories listed
-
Exploration is Harder than Prediction: Cryptographically Separating Reinforcement Learning from Supervised Learning4 Apr 2024 0 repositories listed
-
REACT: Revealing Evolutionary Action Consequence Trajectories for Interpretable Reinforcement Learning4 Apr 2024 0 repositories listed
-
Methodology for Interpretable Reinforcement Learning for Optimizing Mechanical Ventilation3 Apr 2024 0 repositories listed
-
Reinforcement Learning in Categorical Cybernetics3 Apr 2024 0 repositories listed
-
Active Exploration in Bayesian Model-based Reinforcement Learning for Robot Manipulation2 Apr 2024 0 repositories listed
-
Asymptotics of Language Model Alignment2 Apr 2024 0 repositories listed
-
Emergence of Chemotactic Strategies with Multi-Agent Reinforcement Learning2 Apr 2024 0 repositories listed
-
Is Exploration All You Need? Effective Exploration Characteristics for Transfer in Reinforcement Learning2 Apr 2024 0 repositories listed
-
MTLight: Efficient Multi-Task Reinforcement Learning for Traffic Signal Control1 Apr 2024 0 repositories listed
-
Learning Off-policy with Model-based Intrinsic Motivation For Active Online Exploration31 Mar 2024 0 repositories listed
-
Utilizing Maximum Mean Discrepancy Barycenter for Propagating the Uncertainty of Value Functions in Reinforcement Learning31 Mar 2024 0 repositories listed
-
Survey on Large Language Model-Enhanced Reinforcement Learning: Concept, Taxonomy, and Methods30 Mar 2024 0 repositories listed
-
CtRL-Sim: Reactive and Controllable Driving Agents with Offline Reinforcement Learning29 Mar 2024 0 repositories listed
-
Learning Visual Quadrupedal Loco-Manipulation from Demonstrations29 Mar 2024 0 repositories listed
-
Molecular Generative Adversarial Network with Multi-Property Optimization29 Mar 2024 0 repositories listed
-
Nonparametric Bellman Mappings for Reinforcement Learning: Application to Robust Adaptive Filtering29 Mar 2024 0 repositories listed
-
Jointly Training and Pruning CNNs via Learnable Agent Guidance and Alignment28 Mar 2024 0 repositories listed
-
Offline Imitation Learning from Multiple Baselines with Applications to Compiler Optimization28 Mar 2024 0 repositories listed
-
Reinforcement Learning in Agent-Based Market Simulation: Unveiling Realistic Stylized Facts and Behavior28 Mar 2024 0 repositories listed
-
CaT: Constraints as Terminations for Legged Locomotion Reinforcement Learning27 Mar 2024 0 repositories listed
-
FPGA-Based Neural Thrust Controller for UAVs27 Mar 2024 0 repositories listed
-
Image Deraining via Self-supervised Reinforcement Learning27 Mar 2024 0 repositories listed
-
Long and Short-Term Constraints Driven Safe Reinforcement Learning for Autonomous Driving27 Mar 2024 0 repositories listed
-
LORD: Large Models based Opposite Reward Design for Autonomous Driving27 Mar 2024 0 repositories listed
-
Probabilistic Model Checking of Stochastic Reinforcement Learning Policies27 Mar 2024 0 repositories listed
-
Robustness and Visual Explanation for Black Box Image, Video, and ECG Signal Classification with Reinforcement Learning27 Mar 2024 0 repositories listed
-
Safe and Robust Reinforcement Learning: Principles and Practice27 Mar 2024 0 repositories listed
-
Towards Human-Centered Construction Robotics: A Reinforcement Learning-Driven Companion Robot for Contextually Assisting Carpentry Workers27 Mar 2024 0 repositories listed
-
Depending on yourself when you should: Mentoring LLM with RL agents to become the master in cybersecurity games26 Mar 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.