Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 65
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 65 of 152: papers 6,401 to 6,500 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Discovering Command and Control (C2) Channels on Tor and Public Networks Using Reinforcement Learning14 Feb 2024 0 repositories listed
-
Exploiting Estimation Bias in Clipped Double Q-Learning for Continous Control Reinforcement Learning Tasks14 Feb 2024 0 repositories listed
-
How does Your RL Agent Explore? An Optimal Transport Analysis of Occupancy Measure Trajectories14 Feb 2024 0 repositories listed
-
Steady-State Error Compensation for Reinforcement Learning with Quadratic Rewards14 Feb 2024 0 repositories listed
-
Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption14 Feb 2024 0 repositories listed
-
Intelligent Agricultural Management Considering N₂O Emission and Climate Variability with Uncertainties13 Feb 2024 0 repositories listed
-
Optimal Task Assignment and Path Planning using Conflict-Based Search with Precedence and Temporal Constraints13 Feb 2024 0 repositories listed
-
PRDP: Proximal Reward Difference Prediction for Large-Scale Reward Finetuning of Diffusion Models13 Feb 2024 0 repositories listed
-
Provable Traffic Rule Compliance in Safe Reinforcement Learning on the Open Sea13 Feb 2024 0 repositories listed
-
Auxiliary Reward Generation with Transition Distance Representation Learning12 Feb 2024 0 repositories listed
-
IR-Aware ECO Timing Optimization Using Reinforcement Learning12 Feb 2024 0 repositories listed
-
Near-Minimax-Optimal Distributional Reinforcement Learning with a Generative Model12 Feb 2024 0 repositories listed
-
Future Prediction Can be a Strong Evidence of Good History Representation in Partially Observable Environments11 Feb 2024 0 repositories listed
-
Natural Language Reinforcement Learning11 Feb 2024 0 repositories listed
-
Principled Penalty-based Methods for Bilevel Reinforcement Learning and RLHF10 Feb 2024 0 repositories listed
-
ACTER: Diverse and Actionable Counterfactual Sequences for Explaining and Diagnosing RL Policies9 Feb 2024 0 repositories listed
-
High-Precision Geosteering via Reinforcement Learning and Particle Filters9 Feb 2024 0 repositories listed
-
Learn to Teach: Sample-Efficient Privileged Learning for Humanoid Locomotion over Diverse Terrains9 Feb 2024 0 repositories listed
-
RLEEGNet: Integrating Brain-Computer Interfaces with Adaptive AI for Intuitive Responsiveness and High-Accuracy Motor Imagery Classification9 Feb 2024 0 repositories listed
-
Value function interference and greedy action selection in value-based multi-objective reinforcement learning9 Feb 2024 0 repositories listed
-
Differentially Private Deep Model-Based Reinforcement Learning8 Feb 2024 0 repositories listed
-
Federated Offline Reinforcement Learning: Collaborative Single-Policy Coverage Suffices8 Feb 2024 0 repositories listed
-
Real-World Fluid Directed Rigid Body Control via Deep Reinforcement Learning8 Feb 2024 0 repositories listed
-
Scaling Intelligent Agents in Combat Simulations for Wargaming8 Feb 2024 0 repositories listed
-
Context in Public Health for Underserved Communities: A Bayesian Approach to Online Restless Bandits7 Feb 2024 0 repositories listed
-
A Primal-Dual Algorithm for Offline Constrained Reinforcement Learning with Linear MDPs7 Feb 2024 0 repositories listed
-
Code as Reward: Empowering Reinforcement Learning with VLMs7 Feb 2024 0 repositories listed
-
Convergence for Natural Policy Gradient on Infinite-State Queueing MDPs7 Feb 2024 0 repositories listed
-
Learning by Doing: An Online Causal Reinforcement Learning Framework with Causal-Aware Policy7 Feb 2024 0 repositories listed
-
Learning Diverse Policies with Soft Self-Generated Guidance7 Feb 2024 0 repositories listed
-
6 Feb 2024 0 repositories listed Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 8 harvested samples) · 6 pointer-only (licence)
-
No-Regret Reinforcement Learning in Smooth MDPs6 Feb 2024 0 repositories listed
-
Reinforcement Learning from Bagged Reward6 Feb 2024 0 repositories listed
-
Abstracted Trajectory Visualization for Explainability in Reinforcement Learning5 Feb 2024 0 repositories listed
-
Assessing the Impact of Distribution Shift on Reinforcement Learning Performance5 Feb 2024 0 repositories listed
-
Contrastive Diffuser: Planning Towards High Return States via Contrastive Learning5 Feb 2024 0 repositories listed
-
Frugal Actor-Critic: Sample Efficient Off-Policy Deep Reinforcement Learning Using Unique Experiences5 Feb 2024 0 repositories listed
-
DRED: Zero-Shot Transfer in Reinforcement Learning via Data-Regularised Environment Design5 Feb 2024 0 repositories listed
-
Understanding What Affects the Generalization Gap in Visual Reinforcement Learning: Theory and Empirical Evidence5 Feb 2024 0 repositories listed
-
Utility-Based Reinforcement Learning: Unifying Single-objective and Multi-objective Reinforcement Learning5 Feb 2024 0 repositories listed
-
Vision-Language Models Provide Promptable Representations for Reinforcement Learning5 Feb 2024 0 repositories listed
-
A Safe Reinforcement Learning driven Weights-varying Model Predictive Control for Autonomous Vehicle Motion Control4 Feb 2024 0 repositories listed
-
DiffStitch: Boosting Offline Reinforcement Learning with Diffusion-based Trajectory Stitching4 Feb 2024 0 repositories listed
-
Evading Deep Learning-Based Malware Detectors via Obfuscation: A Deep Reinforcement Learning Approach4 Feb 2024 0 repositories listed
-
The Virtues of Pessimism in Inverse Reinforcement Learning4 Feb 2024 0 repositories listed
-
A Survey of Constraint Formulations in Safe Reinforcement Learning3 Feb 2024 0 repositories listed
-
3 Feb 2024 0 repositories listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 5 pointer-only (licence)
-
An Auction-based Marketplace for Model Trading in Federated Learning2 Feb 2024 0 repositories listed
-
Efficient Reinforcement Learning for Routing Jobs in Heterogeneous Queueing Systems2 Feb 2024 0 repositories listed
-
The Political Preferences of LLMs2 Feb 2024 0 repositories listed
-
The RL/LLM Taxonomy Tree: Reviewing Synergies Between Reinforcement Learning and Large Language Models2 Feb 2024 0 repositories listed
-
A Reinforcement Learning Based Controller to Minimize Forces on the Crutches of a Lower-Limb Exoskeleton31 Jan 2024 0 repositories listed
-
Attention Graph for Multi-Robot Social Navigation with Deep Reinforcement Learning31 Jan 2024 0 repositories listed
-
Causal Coordinated Concurrent Reinforcement Learning31 Jan 2024 0 repositories listed
-
Safe Reinforcement Learning-Based Eco-Driving Control for Mixed Traffic Flows With Disturbances31 Jan 2024 0 repositories listed
-
Reinforcement Learning for Versatile, Dynamic, and Robust Bipedal Locomotion Control30 Jan 2024 0 repositories listed
-
Context-Former: Stitching via Latent Conditioned Sequence Modeling29 Jan 2024 0 repositories listed
-
The Indoor-Training Effect: unexpected gains from distribution shifts in the transition function29 Jan 2024 0 repositories listed
-
SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning29 Jan 2024 0 repositories listed
-
Social Interpretable Reinforcement Learning27 Jan 2024 0 repositories listed
-
On the Limitations of Markovian Rewards to Express Multi-Objective, Risk-Sensitive, and Modal Tasks26 Jan 2024 0 repositories listed
-
Constant Stepsize Q-learning: Distributional Convergence, Bias and Extrapolation25 Jan 2024 0 repositories listed
-
Hierarchical Continual Reinforcement Learning via Large Language Model25 Jan 2024 0 repositories listed
-
Learning-based sensing and computing decision for data freshness in edge computing-enabled networks25 Jan 2024 0 repositories listed
-
Learning fast changing slow in spiking neural networks25 Jan 2024 0 repositories listed
-
Sample Efficient Reinforcement Learning by Automatically Learning to Compose Subtasks25 Jan 2024 0 repositories listed
-
Scilab-RL: A software framework for efficient reinforcement learning and cognitive modeling research25 Jan 2024 0 repositories listed
-
A Safe Reinforcement Learning Algorithm for Supervisory Control of Power Plants23 Jan 2024 0 repositories listed
-
Active Inference as a Model of Agency23 Jan 2024 0 repositories listed
-
Building Minimal and Reusable Causal State Abstractions for Reinforcement Learning23 Jan 2024 0 repositories listed
-
Learning safety critics via a non-contractive binary bellman operator23 Jan 2024 0 repositories listed
-
On the Stochastic (Variance-Reduced) Proximal Gradient Method for Regularized Expected Reward Optimization23 Jan 2024 0 repositories listed
-
Towards Socially and Morally Aware RL agent: Reward Design With LLM23 Jan 2024 0 repositories listed
-
Back-stepping Experience Replay with Application to Model-free Reinforcement Learning for a Soft Snake Robot21 Jan 2024 0 repositories listed
-
Constrained Reinforcement Learning for Adaptive Controller Synchronization in Distributed SDN21 Jan 2024 0 repositories listed
-
Large-scale Reinforcement Learning for Diffusion Models20 Jan 2024 0 repositories listed
-
Efficient Training of Generalizable Visuomotor Policies via Control-Aware Augmentation17 Jan 2024 0 repositories listed
-
Crowd-PrefRL: Preference-Based Reward Learning from Crowds17 Jan 2024 0 repositories listed
-
Learning from Imperfect Demonstrations with Self-Supervision for Robotic Manipulation17 Jan 2024 0 repositories listed
-
CycLight: learning traffic signal cooperation with a cycle-level strategy16 Jan 2024 0 repositories listed
-
PRewrite: Prompt Rewriting with Reinforcement Learning16 Jan 2024 0 repositories listed
-
Safe Reinforcement Learning with Free-form Natural Language Constraints and Pre-Trained Language Models15 Jan 2024 0 repositories listed
-
Beyond Sparse Rewards: Enhancing Reinforcement Learning with Language Model Critique in Text Generation14 Jan 2024 0 repositories listed
-
Reinforcement Learning from LLM Feedback to Counteract Goal Misgeneralization14 Jan 2024 0 repositories listed
-
BP(λ): Online Learning via Synthetic Gradients13 Jan 2024 0 repositories listed
-
Discovering Command and Control Channels Using Reinforcement Learning13 Jan 2024 0 repositories listed
-
Mutual Enhancement of Large Language and Reinforcement Learning Models through Bi-Directional Feedback Mechanisms: A Case Study12 Jan 2024 0 repositories listed
-
UNEX-RL: Reinforcing Long-Term Rewards in Multi-Stage Recommender Systems with UNidirectional EXecution12 Jan 2024 0 repositories listed
-
Model-Free Reinforcement Learning for Automated Fluid Administration in Critical Care11 Jan 2024 0 repositories listed
-
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization11 Jan 2024 0 repositories listed
-
Parrot: Pareto-optimal Multi-Reward Reinforcement Learning Framework for Text-to-Image Generation11 Jan 2024 0 repositories listed
-
An Information Theoretic Approach to Interaction-Grounded Learning10 Jan 2024 0 repositories listed
-
Innate-Values-driven Reinforcement Learning based Cooperative Multi-Agent Cognitive Modeling10 Jan 2024 0 repositories listed
-
Reinforcement Learning for Optimizing RAG for Domain Chatbots10 Jan 2024 0 repositories listed
-
Taming "data-hungry" reinforcement learning? Stability in continuous state-action spaces10 Jan 2024 0 repositories listed
-
StarCraftImage: A Dataset For Prototyping Spatial Reasoning Methods For Multi-Agent Environments9 Jan 2024 0 repositories listed
-
Behavioural Cloning in VizDoom8 Jan 2024 0 repositories listed
-
Deep Reinforcement Learning for Multi-Truck Vehicle Routing Problems with Multi-Leg Demand Routes8 Jan 2024 0 repositories listed
-
Long-term Safe Reinforcement Learning with Binary Feedback8 Jan 2024 0 repositories listed
-
NovelGym: A Flexible Ecosystem for Hybrid Planning and Learning Agents Designed for Open Worlds7 Jan 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.