Browse State-of-the-Art › reinforcement-learning › Papers, page 50
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 50 of 135: papers 4,901 to 5,000 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Meta-Reinforcement Learning with Universal Policy Adaptation: Provable Near-Optimality under All-task Optimum Comparator13 Oct 2024 0 repositories listed
-
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning12 Oct 2024 0 repositories listed
-
Reinforcement Learning in Hyperbolic Spaces: Models and Experiments12 Oct 2024 0 repositories listed
-
Towards a Domain-Specific Modelling Environment for Reinforcement Learning12 Oct 2024 0 repositories listed
-
Hierarchical Universal Value Function Approximators11 Oct 2024 0 repositories listed
-
SOLD: Slot Object-Centric Latent Dynamics Models for Relational Manipulation Learning from Pixels11 Oct 2024 0 repositories listed
-
Efficient Reinforcement Learning with Large Language Model Priors10 Oct 2024 0 repositories listed
-
Neuroplastic Expansion in Deep Reinforcement Learning10 Oct 2024 0 repositories listed
-
Offline Hierarchical Reinforcement Learning via Inverse Optimization10 Oct 2024 0 repositories listed
-
Offline Inverse Constrained Reinforcement Learning for Safe-Critical Decision Making in Healthcare10 Oct 2024 0 repositories listed
-
On the grid-sampling limit SDE10 Oct 2024 0 repositories listed
-
Probabilistic Satisfaction of Temporal Logic Constraints in Reinforcement Learning via Adaptive Policy-Switching10 Oct 2024 0 repositories listed
-
On Reward Transferability in Adversarial Inverse Reinforcement Learning: Insights from Random Matrix Theory10 Oct 2024 0 repositories listed
-
Fostering Intrinsic Motivation in Reinforcement Learning with Pretrained Foundation Models9 Oct 2024 0 repositories listed
-
Honesty to Subterfuge: In-Context Reinforcement Learning Can Make Honest Models Reward Hack9 Oct 2024 0 repositories listed
-
MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning9 Oct 2024 0 repositories listed
-
ReinDiffuse: Crafting Physically Plausible Motions with Reinforced Diffusion Model9 Oct 2024 0 repositories listed
-
Transfer Learning for a Class of Cascade Dynamical Systems9 Oct 2024 0 repositories listed
-
Efficient Policy Evaluation with Safety Constraint for Reinforcement Learning8 Oct 2024 0 repositories listed
-
Reinforcement Learning From Imperfect Corrective Actions And Proxy Rewards8 Oct 2024 0 repositories listed
-
Direct Preference Optimization for LLM-Enhanced Recommendation Systems8 Oct 2024 0 repositories listed
-
Solving Multi-Goal Robotic Tasks with Decision Transformer8 Oct 2024 0 repositories listed
-
AlphaRouter: Quantum Circuit Routing with Reinforcement Learning and Tree Search7 Oct 2024 0 repositories listed
-
Designing a Classifier for Active Fire Detection from Multispectral Satellite Imagery Using Neural Architecture Search7 Oct 2024 0 repositories listed
-
Efficient Model-Based Reinforcement Learning Through Optimistic Thompson Sampling7 Oct 2024 0 repositories listed
-
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning7 Oct 2024 0 repositories listed
-
Mastering Chinese Chess AI (Xiangqi) Without Search7 Oct 2024 0 repositories listed
-
Towards using Reinforcement Learning for Scaling and Data Replication in Cloud Systems7 Oct 2024 0 repositories listed
-
AdaMemento: Adaptive Memory-Assisted Policy Optimization for Reinforcement Learning6 Oct 2024 0 repositories listed
-
Data-driven Under Frequency Load Shedding Using Reinforcement Learning6 Oct 2024 0 repositories listed
-
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer6 Oct 2024 0 repositories listed
-
Latent Action Priors for Locomotion with Deep Reinforcement Learning4 Oct 2024 0 repositories listed
-
Training on more Reachable Tasks for Generalisation in Reinforcement Learning4 Oct 2024 0 repositories listed
-
Cross-Embodiment Dexterous Grasping with Reinforcement Learning3 Oct 2024 0 repositories listed
-
Doubly Optimal Policy Evaluation for Reinforcement Learning3 Oct 2024 0 repositories listed
-
Dual Active Learning for Reinforcement Learning from Human Feedback3 Oct 2024 0 repositories listed
-
ComaDICE: Offline Cooperative Multi-Agent Reinforcement Learning with Stationary Distribution Shift Regularization2 Oct 2024 0 repositories listed
-
Finding path and cycle counting formulae in graphs with Deep Reinforcement Learning2 Oct 2024 0 repositories listed
-
Generative Reward Models2 Oct 2024 0 repositories listed
-
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs2 Oct 2024 0 repositories listed
-
LLM-Augmented Symbolic Reinforcement Learning with Landmark-Based Task Decomposition2 Oct 2024 0 repositories listed
-
Sable: a Performant, Efficient and Scalable Sequence Model for MARL2 Oct 2024 0 repositories listed
-
PreND: Enhancing Intrinsic Motivation in Reinforcement Learning through Pre-trained Network Distillation2 Oct 2024 0 repositories listed
-
Realizable Continuous-Space Shields for Safe Reinforcement Learning2 Oct 2024 0 repositories listed
-
RLEF: Grounding Code LLMs in Execution Feedback with Reinforcement Learning2 Oct 2024 0 repositories listed
-
Scalable Reinforcement Learning-based Neural Architecture Search2 Oct 2024 0 repositories listed
-
Contrastive Abstraction for Reinforcement Learning1 Oct 2024 0 repositories listed
-
Energy-Efficient Computation with DVFS using Deep Reinforcement Learning for Multi-Task Systems in Edge Computing28 Sep 2024 0 repositories listed
-
Refutation of Spectral Graph Theory Conjectures with Search Algorithms)27 Sep 2024 0 repositories listed
-
Robust Deep Reinforcement Learning for Volt-VAR Optimization in Active Distribution System under Uncertainty27 Sep 2024 0 repositories listed
-
State-free Reinforcement Learning27 Sep 2024 0 repositories listed
-
A Survey on Neural Architecture Search Based on Reinforcement Learning26 Sep 2024 0 repositories listed
-
Autoregressive Multi-trait Essay Scoring via Reinforcement Learning with Scoring-aware Multiple Rewards26 Sep 2024 0 repositories listed
-
Criticality and Safety Margins for Reinforcement Learning26 Sep 2024 0 repositories listed
-
Inverse Reinforcement Learning with Multiple Planning Horizons26 Sep 2024 0 repositories listed
-
Model-Free versus Model-Based Reinforcement Learning for Fixed-Wing UAV Attitude Control Under Varying Wind Conditions26 Sep 2024 0 repositories listed
-
Navigation in a simplified Urban Flow through Deep Reinforcement Learning26 Sep 2024 0 repositories listed
-
A random measure approach to reinforcement learning in continuous time25 Sep 2024 0 repositories listed
-
A Survey for Deep Reinforcement Learning Based Network Intrusion Detection25 Sep 2024 0 repositories listed
-
Learning Bipedal Walking for Humanoid Robots in Challenging Environments with Obstacle Avoidance25 Sep 2024 0 repositories listed
-
Offline and Distributional Reinforcement Learning for Radio Resource Management25 Sep 2024 0 repositories listed
-
OffRIPP: Offline RL-based Informative Path Planning25 Sep 2024 0 repositories listed
-
Reinforcement Learning for Finite Space Mean-Field Type Games25 Sep 2024 0 repositories listed
-
Revisiting Space Mission Planning: A Reinforcement Learning-Guided Approach for Multi-Debris Rendezvous25 Sep 2024 0 repositories listed
-
Symbolic State Partitioning for Reinforcement Learning25 Sep 2024 0 repositories listed
-
Topological Foundations of Reinforcement Learning25 Sep 2024 0 repositories listed
-
Uncertainty Representations in State-Space Layers for Deep Reinforcement Learning under Partial Observability25 Sep 2024 0 repositories listed
-
A Critical Review of Safe Reinforcement Learning Techniques in Smart Grid Applications24 Sep 2024 0 repositories listed
-
CLSP: High-Fidelity Contrastive Language-State Pre-training for Agent State Representation24 Sep 2024 0 repositories listed
-
The Dark Side of Rich Rewards: Understanding and Mitigating Noise in VLM Rewards24 Sep 2024 0 repositories listed
-
Provably Efficient Exploration in Inverse Constrained Reinforcement Learning24 Sep 2024 0 repositories listed
-
SurgIRL: Towards Life-Long Learning for Surgical Automation by Incremental Reinforcement Learning24 Sep 2024 0 repositories listed
-
Acting for the Right Reasons: Creating Reason-Sensitive Artificial Moral Agents23 Sep 2024 0 repositories listed
-
CANDERE-COACH: Reinforcement Learning from Noisy Feedback23 Sep 2024 0 repositories listed
-
Deep Reinforcement Learning-based Obstacle Avoidance for Robot Movement in Warehouse Environments23 Sep 2024 0 repositories listed
-
Intelligent Routing Algorithm over SDN: Reusable Reinforcement Learning Approach23 Sep 2024 0 repositories listed
-
COSBO: Conservative Offline Simulation-Based Policy Optimization22 Sep 2024 0 repositories listed
-
Scalable Multi-agent Reinforcement Learning for Factory-wide Dynamic Scheduling20 Sep 2024 0 repositories listed
-
SoloParkour: Constrained Reinforcement Learning for Visual Locomotion from Privileged Experience20 Sep 2024 0 repositories listed
-
TACO-RL: Task Aware Prompt Compression Optimization with Reinforcement Learning19 Sep 2024 0 repositories listed
-
The Central Role of the Loss Function in Reinforcement Learning19 Sep 2024 0 repositories listed
-
HARP: Human-Assisted Regrouping with Permutation Invariant Critic for Multi-Agent Reinforcement Learning18 Sep 2024 0 repositories listed
-
Reinforcement Learning with Lie Group Orientations for Robotics18 Sep 2024 0 repositories listed
-
A Reinforcement Learning Environment for Automatic Code Optimization in the MLIR Compiler17 Sep 2024 0 repositories listed
-
Attacking Slicing Network via Side-channel Reinforcement Learning Attack17 Sep 2024 0 repositories listed
-
Linear Jamming Bandits: Learning to Jam 5G-based Coded Communications Systems17 Sep 2024 0 repositories listed
-
On-policy Actor-Critic Reinforcement Learning for Multi-UAV Exploration17 Sep 2024 0 repositories listed
-
A Model-Free Optimal Control Method With Fixed Terminal States and Delay16 Sep 2024 0 repositories listed
-
Disentangling Uncertainty for Safe Social Navigation using Deep Reinforcement Learning16 Sep 2024 0 repositories listed
-
Reinforcement learning-based statistical search strategy for an axion model from flavor16 Sep 2024 0 repositories listed
-
Reinforcement Learning with Quasi-Hyperbolic Discounting16 Sep 2024 0 repositories listed
-
Safety-Oriented Pruning and Interpretation of Reinforcement Learning Policies16 Sep 2024 0 repositories listed
-
SHIRE: Enhancing Sample Efficiency using Human Intuition in REinforcement Learning16 Sep 2024 0 repositories listed
-
Critic as Lyapunov function (CALF): a model-free, stability-ensuring agent15 Sep 2024 0 repositories listed
-
KAN v.s. MLP for Offline Reinforcement Learning15 Sep 2024 0 repositories listed
-
Mitigating Dimensionality in 2D Rectangle Packing Problem under Reinforcement Learning Schema15 Sep 2024 0 repositories listed
-
Planning Transformer: Long-Horizon Offline Reinforcement Learning with Planning Tokens14 Sep 2024 0 repositories listed
-
Curricula for Learning Robust Policies over Factored State Representations in Changing Environments13 Sep 2024 0 repositories listed
-
Quantum-inspired Reinforcement Learning for Synthesizable Drug Design13 Sep 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.