Browse State-of-the-Art › reinforcement-learning › Papers, page 55
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 55 of 135: papers 5,401 to 5,500 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Improving Offline Reinforcement Learning with Inaccurate Simulators7 May 2024 0 repositories listed
-
RACER: Epistemic Risk-Sensitive RL Enables Fast Driving with Fewer Crashes7 May 2024 0 repositories listed
-
TorchDriveEnv: A Reinforcement Learning Benchmark for Autonomous Driving with Reactive, Realistic, and Diverse Non-Playable Characters7 May 2024 0 repositories listed
-
Federated Reinforcement Learning with Constraint Heterogeneity6 May 2024 0 repositories listed
-
Finite-Time Convergence and Sample Complexity of Actor-Critic Multi-Objective Reinforcement Learning5 May 2024 0 repositories listed
-
Safe Reinforcement Learning with Learned Non-Markovian Safety Constraints5 May 2024 0 repositories listed
-
Implicit Safe Set Algorithm for Provably Safe Reinforcement Learning4 May 2024 0 repositories listed
-
Taming Equilibrium Bias in Risk-Sensitive Multi-Agent Reinforcement Learning4 May 2024 0 repositories listed
-
UDUC: An Uncertainty-driven Approach for Learning-based Robust Control4 May 2024 0 repositories listed
-
Learning Robot Soccer from Egocentric Vision with Deep Reinforcement Learning3 May 2024 0 repositories listed
-
Model-based reinforcement learning for protein backbone design3 May 2024 0 repositories listed
-
SocialGFs: Learning Social Gradient Fields for Multi-Agent Reinforcement Learning3 May 2024 0 repositories listed
-
Zero-Sum Positional Differential Games as a Framework for Robust Reinforcement Learning: Deep Q-Learning Approach3 May 2024 0 repositories listed
-
Behavior Imitation for Manipulator Control and Grasping with Deep Reinforcement Learning2 May 2024 0 repositories listed
-
CityLearn v2: Energy-flexible, resilient, occupant-centric, and carbon-aware management of grid-interactive communities2 May 2024 0 repositories listed
-
Constrained Reinforcement Learning Under Model Mismatch2 May 2024 0 repositories listed
-
Goal-conditioned reinforcement learning for ultrasound navigation guidance2 May 2024 0 repositories listed
-
Intelligent Hybrid Resource Allocation in MEC-assisted RAN Slicing Network2 May 2024 0 repositories listed
-
Reinforcement Learning for Edit-Based Non-Autoregressive Neural Machine Translation2 May 2024 0 repositories listed
-
Reinforcement Learning-Guided Semi-Supervised Learning2 May 2024 0 repositories listed
-
Robust Risk-Sensitive Reinforcement Learning with Conditional Value-at-Risk2 May 2024 0 repositories listed
-
Tabular and Deep Reinforcement Learning for Gittins Index2 May 2024 0 repositories listed
-
Towards Interpretable Reinforcement Learning with Constrained Normalizing Flow Policies2 May 2024 0 repositories listed
-
Employing Federated Learning for Training Autonomous HVAC Systems1 May 2024 0 repositories listed
-
MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games1 May 2024 0 repositories listed
-
Portfolio Management using Deep Reinforcement Learning1 May 2024 0 repositories listed
-
Queue-based Eco-Driving at Roundabouts with Reinforcement Learning1 May 2024 0 repositories listed
-
AI, Pluralism, and (Social) Compensation30 Apr 2024 0 repositories listed
-
Control Policy Correction Framework for Reinforcement Learning-based Energy Arbitrage Strategies29 Apr 2024 0 repositories listed
-
Reinforcement Learning Problem Solving with Large Language Models29 Apr 2024 0 repositories listed
-
Resource-rational reinforcement learning and sensorimotor causal states, and resource-rational maximiners29 Apr 2024 0 repositories listed
-
Self-training superconducting neuromorphic circuits using reinforcement learning rules29 Apr 2024 0 repositories listed
-
Verco: Learning Coordinated Verbal Communication for Multi-agent Reinforcement Learning27 Apr 2024 0 repositories listed
-
Knowledge Transfer for Cross-Domain Reinforcement Learning: A Systematic Review26 Apr 2024 0 repositories listed
-
Offline Reinforcement Learning with Behavioral Supervisor Tuning25 Apr 2024 0 repositories listed
-
24 Apr 2024 0 repositories listed Syntology 1 ran (of which 1 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified; the one sample that ran constructed an object rather than computing a result (of 1 harvested sample)
-
GRSN: Gated Recurrent Spiking Neurons for POMDPs and MARL24 Apr 2024 0 repositories listed
-
Cache-Aware Reinforcement Learning in Large-Scale Recommender Systems23 Apr 2024 0 repositories listed
-
Enhancing High-Speed Cruising Performance of Autonomous Vehicles through Integrated Deep Reinforcement Learning Framework23 Apr 2024 0 repositories listed
-
Evolutionary Reinforcement Learning via Cooperative Coevolution23 Apr 2024 0 repositories listed
-
MultiSTOP: Solving Functional Equations with Reinforcement Learning23 Apr 2024 0 repositories listed
-
The Power of Resets in Online Reinforcement Learning23 Apr 2024 0 repositories listed
-
Using deep reinforcement learning to promote sustainable human behaviour on a common pool resource problem23 Apr 2024 0 repositories listed
-
Distributional Black-Box Model Inversion Attack with Multi-Agent Reinforcement Learning22 Apr 2024 0 repositories listed
-
Learning Control Barrier Functions and their application in Reinforcement Learning: A Survey22 Apr 2024 0 repositories listed
-
On the stability of Lipschitz continuous control problems and its application to reinforcement learning20 Apr 2024 0 repositories listed
-
Continuous-time Risk-sensitive Reinforcement Learning via Quadratic Variation Penalty19 Apr 2024 0 repositories listed
-
Mapping Social Choice Theory to RLHF19 Apr 2024 0 repositories listed
-
MM-PhyRLHF: Reinforcement Learning Framework for Multimodal Physics Question-Answering19 Apr 2024 0 repositories listed
-
Random Network Distillation Based Deep Reinforcement Learning for AGV Path Planning19 Apr 2024 0 repositories listed
-
Reinforcement Learning Approach for Integrating Compressed Contexts into Knowledge Graphs19 Apr 2024 0 repositories listed
-
Data-Incremental Continual Offline Reinforcement Learning19 Apr 2024 0 repositories listed
-
Zero-Shot Stitching in Reinforcement Learning using Relative Representations19 Apr 2024 0 repositories listed
-
Actor-Critic Reinforcement Learning with Phased Actor18 Apr 2024 0 repositories listed
-
LTL-Constrained Policy Optimization with Cycle Experience Replay17 Apr 2024 0 repositories listed
-
Function Approximation for Reinforcement Learning Controller for Energy from Spread Waves17 Apr 2024 0 repositories listed
-
Automatic re-calibration of quantum devices by reinforcement learning16 Apr 2024 0 repositories listed
-
Continuous Control Reinforcement Learning: Distributed Distributional DrQ Algorithms16 Apr 2024 0 repositories listed
-
Offline Trajectory Generalization for Offline Reinforcement Learning16 Apr 2024 0 repositories listed
-
Simplex Decomposition for Portfolio Allocation Constraints in Reinforcement Learning16 Apr 2024 0 repositories listed
-
Towards a Research Community in Interpretable Reinforcement Learning: the InterpPol Workshop16 Apr 2024 0 repositories listed
-
Autonomous Path Planning for Intercostal Robotic Ultrasound Imaging Using Reinforcement Learning15 Apr 2024 0 repositories listed
-
Reinforcement Learning from Multi-role Debates as Feedback for Bias Mitigation in LLMs15 Apr 2024 0 repositories listed
-
Effective Reinforcement Learning Based on Structural Information Principles15 Apr 2024 0 repositories listed
-
EyeFormer: Predicting Personalized Scanpaths with Transformer-Guided Reinforcement Learning15 Apr 2024 0 repositories listed
-
Higher Replay Ratio Empowers Sample-Efficient Multi-Agent Reinforcement Learning15 Apr 2024 0 repositories listed
-
On the Effects of Fine-tuning Language Models for Text-Based Reinforcement Learning15 Apr 2024 0 repositories listed
-
The Feasibility of Constrained Reinforcement Learning Algorithms: A Tutorial Study15 Apr 2024 0 repositories listed
-
Knowledgeable Agents by Offline Reinforcement Learning from Large Language Model Rollouts14 Apr 2024 0 repositories listed
-
Model-based Offline Quantum Reinforcement Learning14 Apr 2024 0 repositories listed
-
Active Learning for Control-Oriented Identification of Nonlinear Systems13 Apr 2024 0 repositories listed
-
Safe Reinforcement Learning on the Constraint Manifold: Theory and Applications13 Apr 2024 0 repositories listed
-
Advancing Forest Fire Prevention: Deep Reinforcement Learning for Effective Firebreak Placement12 Apr 2024 0 repositories listed
-
Agile and versatile bipedal robot tracking control through reinforcement learning12 Apr 2024 0 repositories listed
-
RLEMMO: Evolutionary Multimodal Optimization Assisted By Deep Reinforcement Learning12 Apr 2024 0 repositories listed
-
RLHF Deciphered: A Critical Analysis of Reinforcement Learning from Human Feedback for LLMs12 Apr 2024 0 repositories listed
-
SIR-RL: Reinforcement Learning for Optimized Policy Control during Epidemiological Outbreaks in Emerging Market and Developing Economies12 Apr 2024 0 repositories listed
-
Differentially Private Reinforcement Learning with Self-Play11 Apr 2024 0 repositories listed
-
FPGA Divide-and-Conquer Placement using Deep Reinforcement Learning11 Apr 2024 0 repositories listed
-
R2 Indicator and Deep Reinforcement Learning Enhanced Adaptive Multi-Objective Evolutionary Algorithm11 Apr 2024 0 repositories listed
-
UAV-enabled Collaborative Beamforming via Multi-Agent Deep Reinforcement Learning11 Apr 2024 0 repositories listed
-
Dual Ensemble Kalman Filter for Stochastic Optimal Control10 Apr 2024 0 repositories listed
-
Structured Reinforcement Learning for Media Streaming at the Wireless Edge10 Apr 2024 0 repositories listed
-
Deep Reinforcement Learning-Based Approach for a Single Vehicle Persistent Surveillance Problem with Fuel Constraints9 Apr 2024 0 repositories listed
-
Efficient Multi-Task Reinforcement Learning via Task-Specific Action Correction9 Apr 2024 0 repositories listed
-
Generative Pre-Trained Transformer for Symbolic Regression Base In-Context Reinforcement Learning9 Apr 2024 0 repositories listed
-
Graph Reinforcement Learning for Combinatorial Optimization: A Survey and Unifying Perspective9 Apr 2024 0 repositories listed
-
Chiplet Placement Order Exploration Based on Learning to Rank with Graph Representation7 Apr 2024 0 repositories listed
-
Deep Reinforcement Learning Control for Disturbance Rejection in a Nonlinear Dynamic System with Parametric Uncertainty6 Apr 2024 0 repositories listed
-
Structurally Flexible Neural Networks: Evolving the Building Blocks for General Agents6 Apr 2024 0 repositories listed
-
Demonstration Guided Multi-Objective Reinforcement Learning5 Apr 2024 0 repositories listed
-
Enhancing IoT Intelligence: A Transformer-based Reinforcement Learning Methodology5 Apr 2024 0 repositories listed
-
Heterogeneous Multi-Agent Reinforcement Learning for Zero-Shot Scalable Collaboration5 Apr 2024 0 repositories listed
-
Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report5 Apr 2024 0 repositories listed
-
Pixel-wise RL on Diffusion Models: Reinforcement Learning from Rich Feedback5 Apr 2024 0 repositories listed
-
A Reinforcement Learning based Reset Policy for CDCL SAT Solvers4 Apr 2024 0 repositories listed
-
Can Small Language Models Help Large Language Models Reason Better?: LM-Guided Chain-of-Thought4 Apr 2024 0 repositories listed
-
REACT: Revealing Evolutionary Action Consequence Trajectories for Interpretable Reinforcement Learning4 Apr 2024 0 repositories listed
-
Self-organized free-flight arrival for urban air mobility4 Apr 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.