Browse State-of-the-Art › Reinforcement Learning (RL) › Papers, page 74
Reinforcement Learning (RL)
Papers archive 2025-07-28
archive papers tagged: 15,113 · with a code link: 4,749 · where Syntology ran a sample: 1,416 (1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (1,416 of 15,113 tagged: 1,186 with a run with no instrument failure, 230 where every run was a failure of Syntology's instrument)
Page 74 of 152: papers 7,301 to 7,400 of 15,113, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Deep Reinforcement Learning for Multi-user Massive MIMO with Channel Aging14 Feb 2023 0 repositories listed
-
Quantum algorithms applied to satellite mission planning for Earth observation14 Feb 2023 0 repositories listed
-
To Risk or Not to Risk: Learning with Risk Quantification for IoT Task Offloading in UAVs14 Feb 2023 0 repositories listed
-
A Lifetime Extended Energy Management Strategy for Fuel Cell Hybrid Electric Vehicles via Self-Learning Fuzzy Reinforcement Learning13 Feb 2023 0 repositories listed
-
On Modeling Long-Term User Engagement from Stochastic Feedback13 Feb 2023 0 repositories listed
-
Universal Agent Mixtures and the Geometry of Intelligence13 Feb 2023 0 repositories listed
-
Maneuver Decision-Making For Autonomous Air Combat Through Curriculum Learning And Reinforcement Learning With Sparse Rewards12 Feb 2023 0 repositories listed
-
ReMIX: Regret Minimization for Monotonic Value Function Factorization in Multiagent Reinforcement Learning11 Feb 2023 0 repositories listed
-
A Survey on Causal Reinforcement Learning10 Feb 2023 0 repositories listed
-
Low Entropy Communication in Multi-Agent Reinforcement Learning10 Feb 2023 0 repositories listed
-
Towards Minimax Optimality of Model-based Robust Reinforcement Learning10 Feb 2023 0 repositories listed
-
An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning9 Feb 2023 0 repositories listed
-
CLARE: Conservative Model-Based Reward Learning for Offline Inverse Reinforcement Learning9 Feb 2023 0 repositories listed
-
Data Quality-aware Mixed-precision Quantization via Hybrid Reinforcement Learning9 Feb 2023 0 repositories listed
-
Equivariant MuZero9 Feb 2023 0 repositories listed
-
A Near-Optimal Algorithm for Safe Reinforcement Learning Under Instantaneous Hard Constraints8 Feb 2023 0 repositories listed
-
A Scale-Independent Multi-Objective Reinforcement Learning with Convergence Analysis8 Feb 2023 0 repositories listed
-
AISYN: AI-driven Reinforcement Learning-Based Logic Synthesis Framework8 Feb 2023 0 repositories listed
-
Efficient Planning in Combinatorial Action Spaces with Applications to Cooperative Multi-Agent Reinforcement Learning8 Feb 2023 0 repositories listed
-
Near-Optimal Adversarial Reinforcement Learning with Switching Costs8 Feb 2023 0 repositories listed
-
Adaptive Aggregation for Safety-Critical Control7 Feb 2023 0 repositories listed
-
Eliciting User Preferences for Personalized Multi-Objective Decision Making through Comparative Feedback7 Feb 2023 0 repositories listed
-
Ensemble Value Functions for Efficient Exploration in Multi-Agent Reinforcement Learning7 Feb 2023 0 repositories listed
-
Near-Minimax-Optimal Risk-Sensitive Reinforcement Learning with CVaR7 Feb 2023 0 repositories listed
-
Online Reinforcement Learning with Uncertain Episode Lengths7 Feb 2023 0 repositories listed
-
Optimizing Audio Recommendations for the Long-Term: A Reinforcement Learning Perspective7 Feb 2023 0 repositories listed
-
Towards Skilled Population Curriculum for Multi-Agent Reinforcement Learning7 Feb 2023 0 repositories listed
-
Transfer learning for process design with reinforcement learning7 Feb 2023 0 repositories listed
-
A Strong Baseline for Batch Imitation Learning6 Feb 2023 0 repositories listed
-
Arena-Web -- A Web-based Development and Benchmarking Platform for Autonomous Navigation Approaches6 Feb 2023 0 repositories listed
-
DITTO: Offline Imitation Learning with World Models6 Feb 2023 0 repositories listed
-
Holistic Deep-Reinforcement-Learning-based Training of Autonomous Navigation Systems6 Feb 2023 0 repositories listed
-
Offline Learning in Markov Games with General Function Approximation6 Feb 2023 0 repositories listed
-
RLTP: Reinforcement Learning to Pace for Delayed Impression Modeling in Preloaded Ads6 Feb 2023 0 repositories listed
-
State-wise Safe Reinforcement Learning: A Survey6 Feb 2023 0 repositories listed
-
An Online Model-Following Projection Mechanism Using Reinforcement Learning5 Feb 2023 0 repositories listed
-
Open Problems and Modern Solutions for Deep Reinforcement Learning5 Feb 2023 0 repositories listed
-
Offline Minimax Soft-Q-learning Under Realizability and Partial Coverage5 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Traffic Light Control in Intelligent Transportation Systems4 Feb 2023 0 repositories listed
-
Generalization of Deep Reinforcement Learning for Jammer-Resilient Frequency and Power Allocation4 Feb 2023 0 repositories listed
-
Developing Driving Strategies Efficiently: A Skill-Based Hierarchical Reinforcement Learning Approach4 Feb 2023 0 repositories listed
-
Reinforcement Learning in Low-Rank MDPs with Density Features4 Feb 2023 0 repositories listed
-
Reinforcement Learning with History-Dependent Dynamic Contexts4 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Cyber System Defense under Dynamic Adversarial Uncertainties3 Feb 2023 0 repositories listed
-
Deep Reinforcement Learning for Online Error Detection in Cyber-Physical Systems3 Feb 2023 0 repositories listed
-
Reinforcing User Retention in a Billion Scale Short Video Recommender System3 Feb 2023 0 repositories listed
-
2 Feb 2023 0 repositories listed Syntology 5 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 1 where Syntology's instrument failed) · 7 unverified (of 12 harvested samples) · 4 pointer-only (licence)
-
Diversity Through Exclusion (DTE): Niche Identification for Reinforcement Learning through Value-Decomposition2 Feb 2023 0 repositories listed
-
Lower Bounds for Learning in Revealing POMDPs2 Feb 2023 0 repositories listed
-
MARLIN: Soft Actor-Critic based Reinforcement Learning for Congestion Control in Real Networks2 Feb 2023 0 repositories listed
-
Performance Bounds for Policy-Based Average Reward Reinforcement Learning Algorithms2 Feb 2023 0 repositories listed
-
ReLOAD: Reinforcement Learning with Optimistic Ascent-Descent for Last-Iterate Convergence in Constrained MDPs2 Feb 2023 0 repositories listed
-
Bridging Physics-Informed Neural Networks with Reinforcement Learning: Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO)1 Feb 2023 0 repositories listed
-
Combining Deep Reinforcement Learning and Search with Generative Models for Game-Theoretic Opponent Modeling1 Feb 2023 0 repositories listed
-
Multi-zone HVAC Control with Model-Based Deep Reinforcement Learning1 Feb 2023 0 repositories listed
-
Sample Complexity of Kernel-Based Q-Learning1 Feb 2023 0 repositories listed
-
Selective Uncertainty Propagation in Offline RL1 Feb 2023 0 repositories listed
-
Enabling surrogate-assisted evolutionary reinforcement learning via policy embedding31 Jan 2023 0 repositories listed
-
Partitioning Distributed Compute Jobs with Reinforcement Learning and Graph Neural Networks31 Jan 2023 0 repositories listed
-
Towards interpretable quantum machine learning via single-photon quantum walks31 Jan 2023 0 repositories listed
-
Scalable Grid-Aware Dynamic Matching using Deep Reinforcement Learning31 Jan 2023 0 repositories listed
-
Scaling laws for single-agent reinforcement learning31 Jan 2023 0 repositories listed
-
Scheduling Inference Workloads on Distributed Edge Clusters with Reinforcement Learning31 Jan 2023 0 repositories listed
-
Skill Decision Transformer31 Jan 2023 0 repositories listed
-
Hierarchical Programmatic Reinforcement Learning via Learning to Compose Programs30 Jan 2023 0 repositories listed
-
Improved Regret for Efficient Online Reinforcement Learning with Linear Function Approximation30 Jan 2023 0 repositories listed
-
Regret Bounds for Markov Decision Processes with Recursive Optimized Certainty Equivalents30 Jan 2023 0 repositories listed
-
STEEL: Singularity-aware Reinforcement Learning30 Jan 2023 0 repositories listed
-
Transferring Multiple Policies to Hotstart Reinforcement Learning in an Air Compressor Management Problem30 Jan 2023 0 repositories listed
-
V2N Service Scaling with Deep Reinforcement Learning30 Jan 2023 0 repositories listed
-
A Deep Reinforcement Learning Framework for Optimizing Congestion Control in Data Centers29 Jan 2023 0 repositories listed
-
Autonomous Satellite Docking via Adaptive Optimal Output Rregulation: A Reinforcement Learning Approach29 Jan 2023 0 repositories listed
-
Sample Efficient Deep Reinforcement Learning via Local Planning29 Jan 2023 0 repositories listed
-
Beyond Exponentially Fast Mixing in Average-Reward Reinforcement Learning via Multi-Level Monte Carlo Actor-Critic28 Jan 2023 0 repositories listed
-
SaFormer: A Conditional Sequence Modeling Approach to Offline Safe Reinforcement Learning28 Jan 2023 0 repositories listed
-
STEERING: Stein Information Directed Exploration for Model-Based Reinforcement Learning28 Jan 2023 0 repositories listed
-
Towards Learning Rubik's Cube with N-tuple-based Reinforcement Learning28 Jan 2023 0 repositories listed
-
Turbulence control in plane Couette flow using low-dimensional neural ODE-based models and deep reinforcement learning28 Jan 2023 0 repositories listed
-
A Memory Efficient Deep Reinforcement Learning Approach For Snake Game Autonomous Agents27 Jan 2023 0 repositories listed
-
Improving Behavioural Cloning with Positive Unlabeled Learning27 Jan 2023 0 repositories listed
-
Exploring Deep Reinforcement Learning for Holistic Smart Building Control27 Jan 2023 0 repositories listed
-
Generalized Munchausen Reinforcement Learning using Tsallis KL Divergence27 Jan 2023 0 repositories listed
-
Modeling human road crossing decisions as reward maximization with visual perception limitations27 Jan 2023 0 repositories listed
-
Reinforcement Learning from Diverse Human Preferences27 Jan 2023 0 repositories listed
-
Single-Trajectory Distributionally Robust Reinforcement Learning27 Jan 2023 0 repositories listed
-
SNeRL: Semantic-aware Neural Radiance Fields for Reinforcement Learning27 Jan 2023 0 repositories listed
-
Solving Richly Constrained Reinforcement Learning through State Augmentation and Reward Penalties27 Jan 2023 0 repositories listed
-
Communication-Efficient Collaborative Regret Minimization in Multi-Armed Bandits26 Jan 2023 0 repositories listed
-
FedHQL: Federated Heterogeneous Q-Learning26 Jan 2023 0 repositories listed
-
Learning to Generate All Feasible Actions26 Jan 2023 0 repositories listed
-
Model-based Offline Reinforcement Learning with Local Misspecification26 Jan 2023 0 repositories listed
-
On the Global Convergence of Risk-Averse Policy Gradient Methods with Expected Conditional Risk Measures26 Jan 2023 0 repositories listed
-
Certifiably Robust Reinforcement Learning through Model-Based Abstract Interpretation26 Jan 2023 0 repositories listed
-
Principled Reinforcement Learning with Human Feedback from Pairwise or K-wise Comparisons26 Jan 2023 0 repositories listed
-
A Deep Neural Network Algorithm for Linear-Quadratic Portfolio Optimization with MGARCH and Small Transaction Costs25 Jan 2023 0 repositories listed
-
A Novel Deep Reinforcement Learning-based Approach for Enhancing Spectral Efficiency of IRS-assisted Wireless Systems24 Jan 2023 0 repositories listed
-
ASQ-IT: Interactive Explanations for Reinforcement-Learning Agents24 Jan 2023 0 repositories listed
-
AutoCost: Evolving Intrinsic Cost for Zero-violation Reinforcement Learning24 Jan 2023 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.