Browse State-of-the-Art › reinforcement-learning › Papers, page 49
reinforcement-learning
Papers archive 2025-07-28
archive papers tagged: 13,427 · with a code link: 4,119 · where Syntology ran a sample: 1,165 (973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,165 of 13,427 tagged: 973 with a run with no instrument failure, 192 where every run was a failure of Syntology's instrument)
Page 49 of 135: papers 4,801 to 4,900 of 13,427, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Plasticity Loss in Deep Reinforcement Learning: A Survey7 Nov 2024 0 repositories listed
-
Approximate Equivariance in Reinforcement Learning6 Nov 2024 0 repositories listed
-
From Novice to Expert: LLM Agent Policy Optimization via Step-wise Reinforcement Learning6 Nov 2024 0 repositories listed
-
Non-Stationary Learning of Neural Networks with Automatic Soft Parameter Reset6 Nov 2024 0 repositories listed
-
Opportunities of Reinforcement Learning in South Africa's Just Transition6 Nov 2024 0 repositories listed
-
Accelerating Task Generalisation with Multi-Level Skill Hierarchies5 Nov 2024 0 repositories listed
-
Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning5 Nov 2024 0 repositories listed
-
When to Localize? A Risk-Constrained Reinforcement Learning Approach5 Nov 2024 0 repositories listed
-
Learning Hidden Subgoals under Temporal Ordering Constraints in Reinforcement Learning3 Nov 2024 0 repositories listed
-
Teaching Models to Improve on Tape3 Nov 2024 0 repositories listed
-
A Review of Reinforcement Learning in Financial Applications1 Nov 2024 0 repositories listed
-
Effective ML Model Versioning in Edge Networks1 Nov 2024 0 repositories listed
-
Enhancing Adaptive Mixed-Criticality Scheduling with Deep Reinforcement Learning1 Nov 2024 0 repositories listed
-
Safe Imitation Learning-based Optimal Energy Storage Systems Dispatch in Distribution Networks1 Nov 2024 0 repositories listed
-
Statistical Guarantees for Lifelong Reinforcement Learning using PAC-Bayes Theory1 Nov 2024 0 repositories listed
-
Anytime-Constrained Equilibria in Polynomial Time31 Oct 2024 0 repositories listed
-
Compositional Automata Embeddings for Goal-Conditioned Reinforcement Learning31 Oct 2024 0 repositories listed
-
Maximum Entropy Hindsight Experience Replay31 Oct 2024 0 repositories listed
-
Progressive Safeguards for Safe and Model-Agnostic Reinforcement Learning31 Oct 2024 0 repositories listed
-
RL-STaR: Theoretical Analysis of Reinforcement Learning Frameworks for Self-Taught Reasoner31 Oct 2024 0 repositories listed
-
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval30 Oct 2024 0 repositories listed
-
Resource Governance in Networked Systems via Integrated Variational Autoencoders and Reinforcement Learning30 Oct 2024 0 repositories listed
-
Return Augmented Decision Transformer for Off-Dynamics Reinforcement Learning30 Oct 2024 0 repositories listed
-
Self-Driving Car Racing: Application of Deep Reinforcement Learning30 Oct 2024 0 repositories listed
-
Stepping Out of the Shadows: Reinforcement Learning in Shadow Mode30 Oct 2024 0 repositories listed
-
Hindsight Experience Replay Accelerates Proximal Policy Optimization29 Oct 2024 0 repositories listed
-
PrefPaint: Aligning Image Inpainting Diffusion Model with Human Preference29 Oct 2024 0 repositories listed
-
Solving Minimum-Cost Reach Avoid using Reinforcement Learning29 Oct 2024 0 repositories listed
-
A Multi-Agent Reinforcement Learning Testbed for Cognitive Radio Applications28 Oct 2024 0 repositories listed
-
Active Legibility in Multiagent Reinforcement Learning28 Oct 2024 0 repositories listed
-
Adversarial Constrained Policy Optimization: Improving Constrained Reinforcement Learning by Adapting Budgets28 Oct 2024 0 repositories listed
-
Dual-Agent Deep Reinforcement Learning for Dynamic Pricing and Replenishment28 Oct 2024 0 repositories listed
-
Exploring reinforcement learning for incident response in autonomous military vehicles28 Oct 2024 0 repositories listed
-
Offline Reinforcement Learning With Combinatorial Action Spaces28 Oct 2024 0 repositories listed
-
Quantum Reinforcement Learning-Based Two-Stage Unit Commitment Framework for Enhanced Power Systems Robustness28 Oct 2024 0 repositories listed
-
Video to Video Generative Adversarial Network for Few-shot Learning Based on Policy Gradient28 Oct 2024 0 repositories listed
-
Efficient Diversity-based Experience Replay for Deep Reinforcement Learning27 Oct 2024 0 repositories listed
-
Overcoming the Sim-to-Real Gap: Leveraging Simulation to Learn to Explore for Real-World RL26 Oct 2024 0 repositories listed
-
Uncertainty-Penalized Direct Preference Optimization26 Oct 2024 0 repositories listed
-
Evolving choice hysteresis in reinforcement learning: comparing the adaptive value of positivity bias and gradual perseveration25 Oct 2024 0 repositories listed
-
MILES: Making Imitation Learning Easy with Self-Supervision25 Oct 2024 0 repositories listed
-
Multi-Agent Reinforcement Learning with Selective State-Space Models25 Oct 2024 0 repositories listed
-
Offline-to-Online Multi-Agent Reinforcement Learning with Offline Value Function Memory and Sequential Exploration25 Oct 2024 0 repositories listed
-
On-Robot Reinforcement Learning with Goal-Contrastive Rewards25 Oct 2024 0 repositories listed
-
Provably Adaptive Average Reward Reinforcement Learning for Metric Spaces25 Oct 2024 0 repositories listed
-
Random Policy Enables In-Context Reinforcement Learning within Trust Horizons25 Oct 2024 0 repositories listed
-
PointPatchRL -- Masked Reconstruction Improves Reinforcement Learning on Point Clouds24 Oct 2024 0 repositories listed
-
Reinforcement Learning the Chromatic Symmetric Function24 Oct 2024 0 repositories listed
-
SAMG: State-Action-Aware Offline-to-Online Reinforcement Learning with Offline Model Guidance24 Oct 2024 0 repositories listed
-
The Hive Mind is a Single Reinforcement Learning Agent23 Oct 2024 0 repositories listed
-
Multimodal Information Bottleneck for Deep Reinforcement Learning with Multiple Sensors23 Oct 2024 0 repositories listed
-
Reinforcement Learning under Latent Dynamics: Toward Statistical and Algorithmic Modularity23 Oct 2024 0 repositories listed
-
SPIRE: Synergistic Planning, Imitation, and Reinforcement Learning for Long-Horizon Manipulation23 Oct 2024 0 repositories listed
-
DROP: Distributional and Regular Optimism and Pessimism for Reinforcement Learning22 Oct 2024 0 repositories listed
-
Episodic Future Thinking Mechanism for Multi-agent Reinforcement Learning22 Oct 2024 0 repositories listed
-
Large Language Models are In-context Preference Learners22 Oct 2024 0 repositories listed
-
Hierarchical Multi-agent Reinforcement Learning for Cyber Network Defense22 Oct 2024 0 repositories listed
-
Multi-Modal Transformer and Reinforcement Learning-based Beam Management22 Oct 2024 0 repositories listed
-
QuasiNav: Asymmetric Cost-Aware Navigation Planning with Constrained Quasimetric Reinforcement Learning22 Oct 2024 0 repositories listed
-
Curriculum Reinforcement Learning for Complex Reward Functions22 Oct 2024 0 repositories listed
-
A Plug-and-Play Fully On-the-Job Real-Time Reinforcement Learning Algorithm for a Direct-Drive Tandem-Wing Experiment Platforms Under Multiple Random Operating Conditions21 Oct 2024 0 repositories listed
-
Advancements in Electric Vehicle Charging Optimization: A Survey of Reinforcement Learning Approaches21 Oct 2024 0 repositories listed
-
AttentionPainter: An Efficient and Adaptive Stroke Predictor for Scene Painting21 Oct 2024 0 repositories listed
-
Offline reinforcement learning for job-shop scheduling problems21 Oct 2024 0 repositories listed
-
RGMDT: Return-Gap-Minimizing Decision Tree Extraction in Non-Euclidean Metric Space21 Oct 2024 0 repositories listed
-
Understanding and Alleviating Memory Consumption in RLHF for LLMs21 Oct 2024 0 repositories listed
-
AssemblyComplete: 3D Combinatorial Construction with Deep Reinforcement Learning20 Oct 2024 0 repositories listed
-
A Novel Reinforcement Learning Model for Post-Incident Malware Investigations19 Oct 2024 0 repositories listed
-
GNNRL-Smoothing: A Prior-Free Reinforcement Learning Model for Mesh Smoothing19 Oct 2024 0 repositories listed
-
Hierarchical Reinforced Trader (HRT): A Bi-Level Approach for Optimizing Stock Selection and Execution19 Oct 2024 0 repositories listed
-
Semantic Information G Theory for Range Control with Tradeoff between Purposiveness and Efficiency19 Oct 2024 0 repositories listed
-
Harnessing Causality in Reinforcement Learning With Bagged Decision Times18 Oct 2024 0 repositories listed
-
Interpretable end-to-end Neurosymbolic Reinforcement Learning agents18 Oct 2024 0 repositories listed
-
Inverse Reinforcement Learning from Non-Stationary Learning Agents18 Oct 2024 0 repositories listed
-
Online Reinforcement Learning with Passive Memory18 Oct 2024 0 repositories listed
-
Reinforcement Learning in Non-Markov Market-Making18 Oct 2024 0 repositories listed
-
Transfer Reinforcement Learning in Heterogeneous Action Spaces using Subgoal Mapping18 Oct 2024 0 repositories listed
-
Utilizing Large Language Models for Event Deconstruction to Enhance Multimodal Aspect-Based Sentiment Analysis18 Oct 2024 0 repositories listed
-
Adversarial Inception Backdoor Attacks against Reinforcement Learning17 Oct 2024 0 repositories listed
-
Approximating Auction Equilibria with Reinforcement Learning17 Oct 2024 0 repositories listed
-
Guided Reinforcement Learning for Robust Multi-Contact Loco-Manipulation17 Oct 2024 0 repositories listed
-
Is Prior-Free Black-Box Non-Stationary Reinforcement Learning Feasible?17 Oct 2024 0 repositories listed
-
Rethinking Optimal Transport in Offline Reinforcement Learning17 Oct 2024 0 repositories listed
-
Dynamic Learning Rate for Deep Reinforcement Learning: A Bandit Approach16 Oct 2024 0 repositories listed
-
GAN Based Top-Down View Synthesis in Reinforcement Learning Environments16 Oct 2024 0 repositories listed
-
Insights from the Inverse: Reconstructing LLM Training Goals Through Inverse RL16 Oct 2024 0 repositories listed
-
Spectrum Sharing using Deep Reinforcement Learning in Vehicular Networks16 Oct 2024 0 repositories listed
-
Advanced Persistent Threats (APT) Attribution Using Deep Reinforcement Learning15 Oct 2024 0 repositories listed
-
Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning15 Oct 2024 0 repositories listed
-
DODT: Enhanced Online Decision Transformer Learning through Dreamer's Actor-Critic Trajectory Forecasting15 Oct 2024 0 repositories listed
-
Multi-objective Reinforcement Learning: A Tool for Pluralistic Alignment15 Oct 2024 0 repositories listed
-
Physical Informed-Inspired Deep Reinforcement Learning Based Bi-Level Programming for Microgrid Scheduling15 Oct 2024 0 repositories listed
-
Solving The Dynamic Volatility Fitting Problem: A Deep Reinforcement Learning Approach15 Oct 2024 0 repositories listed
-
Burning RED: Unlocking Subtask-Driven Reinforcement Learning and Risk-Awareness in Average-Reward Markov Decision Processes14 Oct 2024 0 repositories listed
-
Compositional Shielding and Reinforcement Learning for Multi-Agent Systems14 Oct 2024 0 repositories listed
-
Diversity-Aware Reinforcement Learning for de novo Drug Design14 Oct 2024 0 repositories listed
-
Enhancing Robustness in Deep Reinforcement Learning: A Lyapunov Exponent Approach14 Oct 2024 0 repositories listed
-
QE-EBM: Using Quality Estimators as Energy Loss for Machine Translation14 Oct 2024 0 repositories listed
Syntology lines on 2 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.