Browse State-of-the-Art › Reinforcement Learning › Papers, page 50
Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 13,178 · with a code link: 4,183 · where Syntology ran a sample: 1,175 (988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,175 of 13,178 tagged: 988 with a run with no instrument failure, 187 where every run was a failure of Syntology's instrument)
Page 50 of 132: papers 4,901 to 5,000 of 13,178, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Video to Video Generative Adversarial Network for Few-shot Learning Based on Policy Gradient28 Oct 2024 0 repositories listed
-
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning27 Oct 2024 0 repositories listed
-
Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RL26 Oct 2024 0 repositories listed
-
Uncertainty-Penalized Direct Preference Optimization26 Oct 2024 0 repositories listed
-
Evolving choice hysteresis in reinforcement learning: comparing the adaptive value of positivity bias and gradual perseveration25 Oct 2024 0 repositories listed
-
MILES: Making Imitation Learning Easy with Self-Supervision25 Oct 2024 0 repositories listed
-
Multi-Agent Reinforcement Learning with Selective State-Space Models25 Oct 2024 0 repositories listed
-
Offline-to-Online Multi-Agent Reinforcement Learning with Offline Value Function Memory and Sequential Exploration25 Oct 2024 0 repositories listed
-
On-Robot Reinforcement Learning with Goal-Contrastive Rewards25 Oct 2024 0 repositories listed
-
Provably Adaptive Average Reward Reinforcement Learning for Metric Spaces25 Oct 2024 0 repositories listed
-
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons25 Oct 2024 0 repositories listed
-
PointPatchRL -- Masked Reconstruction Improves Reinforcement Learning on Point Clouds24 Oct 2024 0 repositories listed
-
Reinforcement Learning the Chromatic Symmetric Function24 Oct 2024 0 repositories listed
-
SAMG: State-Action-Aware Offline-to-Online Reinforcement Learning with Offline Model Guidance24 Oct 2024 0 repositories listed
-
The Hive Mind is a Single Reinforcement Learning Agent23 Oct 2024 0 repositories listed
-
Multimodal Information Bottleneck for Deep Reinforcement Learning with Multiple Sensors23 Oct 2024 0 repositories listed
-
Reinforcement Learning under Latent Dynamics: Toward Statistical and Algorithmic Modularity23 Oct 2024 0 repositories listed
-
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation23 Oct 2024 0 repositories listed
-
DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning22 Oct 2024 0 repositories listed
-
Episodic Future Thinking Mechanism for Multi-agent Reinforcement Learning22 Oct 2024 0 repositories listed
-
Large Language Models are In-context Preference Learners22 Oct 2024 0 repositories listed
-
Hierarchical Multi-agent Reinforcement Learning for Cyber Network Defense22 Oct 2024 0 repositories listed
-
Multi-Modal Transformer and Reinforcement Learning-based Beam Management22 Oct 2024 0 repositories listed
-
QuasiNav: Asymmetric Cost-Aware Navigation Planning with Constrained Quasimetric Reinforcement Learning22 Oct 2024 0 repositories listed
-
Curriculum Reinforcement Learning for Complex Reward Functions22 Oct 2024 0 repositories listed
-
A Plug-and-Play Fully On-the-Job Real-Time Reinforcement Learning Algorithm for a Direct-Drive Tandem-Wing Experiment Platforms Under Multiple Random Operating Conditions21 Oct 2024 0 repositories listed
-
Advancements in Electric Vehicle Charging Optimization: A Survey of Reinforcement Learning Approaches21 Oct 2024 0 repositories listed
-
AttentionPainter: An Efficient and Adaptive Stroke Predictor for Scene Painting21 Oct 2024 0 repositories listed
-
Offline reinforcement learning for job-shop scheduling problems21 Oct 2024 0 repositories listed
-
RGMDT: Return-Gap-Minimizing Decision Tree Extraction in Non-Euclidean Metric Space21 Oct 2024 0 repositories listed
-
Understanding and Alleviating Memory Consumption in RLHF for LLMs21 Oct 2024 0 repositories listed
-
AssemblyComplete: 3D Combinatorial Construction with Deep Reinforcement Learning20 Oct 2024 0 repositories listed
-
A Novel Reinforcement Learning Model for Post-Incident Malware Investigations19 Oct 2024 0 repositories listed
-
GNNRL-Smoothing: A Prior-Free Reinforcement Learning Model for Mesh Smoothing19 Oct 2024 0 repositories listed
-
Hierarchical Reinforced Trader (HRT): A Bi-Level Approach for Optimizing Stock Selection and Execution19 Oct 2024 0 repositories listed
-
Semantic Information G Theory for Range Control with Tradeoff between Purposiveness and Efficiency19 Oct 2024 0 repositories listed
-
Harnessing Causality in Reinforcement Learning With Bagged Decision Times18 Oct 2024 0 repositories listed
-
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents18 Oct 2024 0 repositories listed
-
Inverse Reinforcement Learning from Non-Stationary Learning Agents18 Oct 2024 0 repositories listed
-
Online Reinforcement Learning with Passive Memory18 Oct 2024 0 repositories listed
-
Reinforcement Learning in Non-Markov Market-Making18 Oct 2024 0 repositories listed
-
Transfer Reinforcement Learning in Heterogeneous Action Spaces using Subgoal Mapping18 Oct 2024 0 repositories listed
-
Utilizing Large Language Models for Event Deconstruction to Enhance Multimodal Aspect-Based Sentiment Analysis18 Oct 2024 0 repositories listed
-
Adversarial Inception Backdoor Attacks against Reinforcement Learning17 Oct 2024 0 repositories listed
-
Approximating Auction Equilibria with Reinforcement Learning17 Oct 2024 0 repositories listed
-
Guided Reinforcement Learning for Robust Multi-Contact Loco-Manipulation17 Oct 2024 0 repositories listed
-
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?17 Oct 2024 0 repositories listed
-
Rethinking Optimal Transport in Offline Reinforcement Learning17 Oct 2024 0 repositories listed
-
Dynamic Learning Rate for Deep Reinforcement Learning: A Bandit Approach16 Oct 2024 0 repositories listed
-
GAN Based Top-Down View Synthesis in Reinforcement Learning Environments16 Oct 2024 0 repositories listed
-
Insights from the Inverse: Reconstructing LLM Training Goals Through Inverse RL16 Oct 2024 0 repositories listed
-
Spectrum Sharing using Deep Reinforcement Learning in Vehicular Networks16 Oct 2024 0 repositories listed
-
Advanced Persistent Threats (APT) Attribution Using Deep Reinforcement Learning15 Oct 2024 0 repositories listed
-
Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning15 Oct 2024 0 repositories listed
-
DODT: Enhanced Online Decision Transformer Learning through Dreamer's Actor-Critic Trajectory Forecasting15 Oct 2024 0 repositories listed
-
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment15 Oct 2024 0 repositories listed
-
Physical Informed-Inspired Deep Reinforcement Learning Based Bi-Level Programming for Microgrid Scheduling15 Oct 2024 0 repositories listed
-
Solving The Dynamic Volatility Fitting Problem: A Deep Reinforcement Learning Approach15 Oct 2024 0 repositories listed
-
Burning RED: Unlocking Subtask-Driven Reinforcement Learning and Risk-Awareness in Average-Reward Markov Decision Processes14 Oct 2024 0 repositories listed
-
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems14 Oct 2024 0 repositories listed
-
Diversity-Aware Reinforcement Learning for de novo Drug Design14 Oct 2024 0 repositories listed
-
Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent Approach14 Oct 2024 0 repositories listed
-
QE-EBM: Using Quality Estimators as Energy Loss for Machine Translation14 Oct 2024 0 repositories listed
-
Meta-Reinforcement Learning with Universal Policy Adaptation: Provable Near-Optimality under All-task Optimum Comparator13 Oct 2024 0 repositories listed
-
ActSafe: Active Exploration with Safety Constraints for Reinforcement Learning12 Oct 2024 0 repositories listed
-
Reinforcement Learning in Hyperbolic Spaces: Models and Experiments12 Oct 2024 0 repositories listed
-
Towards a Domain-Specific Modelling Environment for Reinforcement Learning12 Oct 2024 0 repositories listed
-
Hierarchical Universal Value Function Approximators11 Oct 2024 0 repositories listed
-
SOLD: Slot Object-Centric Latent Dynamics Models for Relational Manipulation Learning from Pixels11 Oct 2024 0 repositories listed
-
Efficient Reinforcement Learning with Large Language Model Priors10 Oct 2024 0 repositories listed
-
Neuroplastic Expansion in Deep Reinforcement Learning10 Oct 2024 0 repositories listed
-
Offline Hierarchical Reinforcement Learning via Inverse Optimization10 Oct 2024 0 repositories listed
-
Offline Inverse Constrained Reinforcement Learning for Safe-Critical Decision Making in Healthcare10 Oct 2024 0 repositories listed
-
On the grid-sampling limit SDE10 Oct 2024 0 repositories listed
-
Probabilistic Satisfaction of Temporal Logic Constraints in Reinforcement Learning via Adaptive Policy-Switching10 Oct 2024 0 repositories listed
-
On Reward Transferability in Adversarial Inverse Reinforcement Learning: Insights from Random Matrix Theory10 Oct 2024 0 repositories listed
-
Fostering Intrinsic Motivation in Reinforcement Learning with Pretrained Foundation Models9 Oct 2024 0 repositories listed
-
Honesty to Subterfuge: In-Context Reinforcement Learning Can Make Honest Models Reward Hack9 Oct 2024 0 repositories listed
-
MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning9 Oct 2024 0 repositories listed
-
ReinDiffuse: Crafting Physically Plausible Motions with Reinforced Diffusion Model9 Oct 2024 0 repositories listed
-
Transfer Learning for a Class of Cascade Dynamical Systems9 Oct 2024 0 repositories listed
-
Efficient Policy Evaluation with Safety Constraint for Reinforcement Learning8 Oct 2024 0 repositories listed
-
Reinforcement Learning From Imperfect Corrective Actions And Proxy Rewards8 Oct 2024 0 repositories listed
-
Direct Preference Optimization for LLM-Enhanced Recommendation Systems8 Oct 2024 0 repositories listed
-
Solving Multi-Goal Robotic Tasks with Decision Transformer8 Oct 2024 0 repositories listed
-
AlphaRouter: Quantum Circuit Routing with Reinforcement Learning and Tree Search7 Oct 2024 0 repositories listed
-
Designing a Classifier for Active Fire Detection from Multispectral Satellite Imagery Using Neural Architecture Search7 Oct 2024 0 repositories listed
-
Efficient Model-Based Reinforcement Learning Through Optimistic Thompson Sampling7 Oct 2024 0 repositories listed
-
HERO: Human-Feedback Efficient Reinforcement Learning for Online Diffusion Model Finetuning7 Oct 2024 0 repositories listed
-
Mastering Chinese Chess AI (Xiangqi) Without Search7 Oct 2024 0 repositories listed
-
Towards using Reinforcement Learning for Scaling and Data Replication in Cloud Systems7 Oct 2024 0 repositories listed
-
AdaMemento: Adaptive Memory-Assisted Policy Optimization for Reinforcement Learning6 Oct 2024 0 repositories listed
-
Data-driven Under Frequency Load Shedding Using Reinforcement Learning6 Oct 2024 0 repositories listed
-
Toward Debugging Deep Reinforcement Learning Programs with RLExplorer6 Oct 2024 0 repositories listed
-
Latent Action Priors for Locomotion with Deep Reinforcement Learning4 Oct 2024 0 repositories listed
-
Training on more Reachable Tasks for Generalisation in Reinforcement Learning4 Oct 2024 0 repositories listed
-
Cross-Embodiment Dexterous Grasping with Reinforcement Learning3 Oct 2024 0 repositories listed
-
Doubly Optimal Policy Evaluation for Reinforcement Learning3 Oct 2024 0 repositories listed
-
Dual Active Learning for Reinforcement Learning from Human Feedback3 Oct 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.