Browse State-of-the-Art › Offline RL › Papers, page 7
Offline RL
Papers archive 2025-07-28
archive papers tagged: 755 · with a code link: 310 · where Syntology ran a sample: 164 (139 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (164 of 755 tagged: 139 with a run with no instrument failure, 25 where every run was a failure of Syntology's instrument)
Page 7 of 8: papers 601 to 700 of 755, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Offline Reinforcement Learning with Closed-Form Policy Improvement Operators29 Nov 2022 0 repositories listed
-
Is Conditional Generative Modeling all you need for Decision-Making?28 Nov 2022 0 repositories listed
-
Offline Q-Learning on Diverse Multi-Task Data Both Scales And Generalizes28 Nov 2022 0 repositories listed
-
State-Aware Proximal Pessimistic Algorithms for Offline Reinforcement Learning28 Nov 2022 0 repositories listed
-
Domain Generalization for Robust Model-Based Offline Reinforcement Learning27 Nov 2022 0 repositories listed
-
On Instance-Dependent Bounds for Offline Reinforcement Learning with Linear Function Approximation23 Nov 2022 0 repositories listed
-
Contextual Transformer for Offline Meta Reinforcement Learning15 Nov 2022 0 repositories listed
-
Offline Reinforcement Learning with Adaptive Behavior Regularization15 Nov 2022 0 repositories listed
-
Leveraging Offline Data in Online Reinforcement Learning9 Nov 2022 0 repositories listed
-
ARMOR: A Model-based Framework for Improving Arbitrary Baseline Policies with Offline Data8 Nov 2022 0 repositories listed
-
Wall Street Tree Search: Risk-Aware Planning for Offline Reinforcement Learning6 Nov 2022 0 repositories listed
-
Contrastive Value Learning: Implicit Models for Simple Offline RL3 Nov 2022 0 repositories listed
-
Oracle Inequalities for Model Selection in Offline Reinforcement Learning3 Nov 2022 0 repositories listed
-
Offline RL With Realistic Datasets: Heteroskedasticity and Support Constraints2 Nov 2022 0 repositories listed
-
Optimal Conservative Offline RL with General Function Approximation via Augmented Lagrangian1 Nov 2022 0 repositories listed
-
Implicit Offline Reinforcement Learning via Supervised Learning21 Oct 2022 0 repositories listed
-
Boosting Offline Reinforcement Learning via Data Rebalancing17 Oct 2022 0 repositories listed
-
Data-Efficient Pipeline for Offline Reinforcement Learning with Limited Data16 Oct 2022 0 repositories listed
-
The Role of Coverage in Online Reinforcement Learning9 Oct 2022 0 repositories listed
-
Offline Reinforcement Learning with Differentiable Function Approximation is Provably Efficient3 Oct 2022 0 repositories listed
-
Offline Reinforcement Learning with Instrumental Variables in Confounded Markov Decision Processes18 Sep 2022 0 repositories listed
-
Can Offline Reinforcement Learning Help Natural Language Understanding?15 Sep 2022 0 repositories listed
-
Distributionally Robust Offline Reinforcement Learning with Linear Function Approximation14 Sep 2022 0 repositories listed
-
Task-Agnostic Learning to Accomplish New Tasks9 Sep 2022 0 repositories listed
-
Dialogue Evaluation with Offline Reinforcement Learning2 Sep 2022 0 repositories listed
-
Strategic Decision-Making in the Presence of Information Asymmetry: Provably Efficient RL with Algorithmic Instruments23 Aug 2022 0 repositories listed
-
Distributionally Robust Model-Based Offline Reinforcement Learning with Near-Optimal Sample Complexity11 Aug 2022 0 repositories listed
-
Offline Reinforcement Learning at Multiple Frequencies26 Jul 2022 0 repositories listed
-
BCRLSP: An Offline Reinforcement Learning Framework for Sequential Targeted Promotion16 Jul 2022 0 repositories listed
-
GriddlyJS: A Web IDE for Reinforcement Learning13 Jul 2022 0 repositories listed
-
An Empirical Study of Implicit Regularization in Deep Offline RL5 Jul 2022 0 repositories listed
-
Offline RL Policies Should be Trained to be Adaptive5 Jul 2022 0 repositories listed
-
Prompting Decision Transformer for Few-Shot Policy Generalization27 Jun 2022 0 repositories listed
-
A Survey on Model-based Reinforcement Learning19 Jun 2022 0 repositories listed
-
Contrastive Learning as Goal-Conditioned Reinforcement Learning15 Jun 2022 0 repositories listed
-
Provable Benefit of Multitask Representation Learning in Reinforcement Learning13 Jun 2022 0 repositories listed
-
Provably Efficient Offline Reinforcement Learning with Trajectory-Wise Reward13 Jun 2022 0 repositories listed
-
Federated Offline Reinforcement Learning11 Jun 2022 0 repositories listed
-
Large-Scale Retrieval for Reinforcement Learning10 Jun 2022 0 repositories listed
-
On the Role of Discount Factor in Offline Reinforcement Learning7 Jun 2022 0 repositories listed
-
Offline Reinforcement Learning with Causal Structured World Models3 Jun 2022 0 repositories listed
-
Offline Reinforcement Learning with Differential Privacy2 Jun 2022 0 repositories listed
-
Know Your Boundaries: The Necessity of Explicit Behavioral Cloning in Offline RL1 Jun 2022 0 repositories listed
-
Model Generation with Provable Coverability for Offline Reinforcement Learning1 Jun 2022 0 repositories listed
-
Nearly Minimax Optimal Offline Reinforcement Learning with Linear Function Approximation: Single-Agent MDP and Markov Game31 May 2022 0 repositories listed
-
You Can't Count on Luck: Why Decision Transformers and RvS Fail in Stochastic Environments31 May 2022 0 repositories listed
-
Pessimism in the Face of Confounders: Provably Efficient Offline Reinforcement Learning in Partially Observable Markov Decision Processes26 May 2022 0 repositories listed
-
How to Spend Your Robot Time: Bridging Kickstarting and Offline Reinforcement Learning for Vision-based Robotic Manipulation6 May 2022 0 repositories listed
-
Pessimism meets VCG: Learning Dynamic Mechanism Design via Offline Reinforcement Learning5 May 2022 0 repositories listed
-
Towards Flexible Inference in Sequential Decision Problems via Bidirectional Transformers28 Apr 2022 0 repositories listed
-
Learning Value Functions from Undirected State-only Experience26 Apr 2022 0 repositories listed
-
When Should We Prefer Offline Reinforcement Learning Over Behavioral Cloning?12 Apr 2022 0 repositories listed
-
Settling the Sample Complexity of Model-Based Offline Reinforcement Learning11 Apr 2022 0 repositories listed
-
A Conservative Q-Learning approach for handling distribution shift in sepsis treatment strategies25 Mar 2022 0 repositories listed
-
Offline Reinforcement Learning Under Value and Density-Ratio Realizability: The Power of Gaps25 Mar 2022 0 repositories listed
-
Bellman Residual Orthogonalization for Offline Reinforcement Learning24 Mar 2022 0 repositories listed
-
Optimizing Trajectories for Highway Driving with Offline Reinforcement Learning21 Mar 2022 0 repositories listed
-
DARA: Dynamics-Aware Reward Augmentation in Offline Reinforcement Learning13 Mar 2022 0 repositories listed
-
Reliable validation of Reinforcement Learning Benchmarks2 Mar 2022 0 repositories listed
-
Pessimistic Q-Learning for Offline Reinforcement Learning: Towards Optimal Sample Complexity28 Feb 2022 0 repositories listed
-
Settling the Communication Complexity for Distributed Offline Reinforcement Learning10 Feb 2022 0 repositories listed
-
Offline Reinforcement Learning with Realizability and Single-policy Concentrability9 Feb 2022 0 repositories listed
-
Transferred Q-learning9 Feb 2022 0 repositories listed
-
How to Leverage Unlabeled Data in Offline Reinforcement Learning3 Feb 2022 0 repositories listed
-
The Challenges of Exploration for Offline Reinforcement Learning27 Jan 2022 0 repositories listed
-
Offline Reinforcement Learning for Road Traffic Control7 Jan 2022 0 repositories listed
-
Importance of Empirical Sample Complexity Analysis for Offline Reinforcement Learning31 Dec 2021 0 repositories listed
-
Single-Shot Pruning for Offline Reinforcement Learning31 Dec 2021 0 repositories listed
-
A Validation Tool for Designing Reinforcement Learning Environments10 Dec 2021 0 repositories listed
-
DR3: Value-Based Deep Reinforcement Learning Requires Explicit Regularization9 Dec 2021 0 repositories listed
-
Curriculum Offline Imitating Learning1 Dec 2021 0 repositories listed
-
Improving Zero-shot Generalization in Offline Reinforcement Learning using Generalized Similarity Functions29 Nov 2021 0 repositories listed
-
UMBRELLA: Uncertainty-Aware Model-Based Offline Reinforcement Learning Leveraging Planning22 Nov 2021 0 repositories listed
-
Offline Reinforcement Learning: Fundamental Barriers for Value Function Approximation21 Nov 2021 0 repositories listed
-
A Survey of Zero-shot Generalisation in Deep Reinforcement Learning18 Nov 2021 0 repositories listed
-
2 Nov 2021 0 repositories listed
-
Towards Instance-Optimal Offline Reinforcement Learning with Pessimism17 Oct 2021 0 repositories listed
-
Value Penalized Q-Learning for Recommender Systems15 Oct 2021 0 repositories listed
-
Representation Learning for Online and Offline RL in Low-rank MDPs9 Oct 2021 0 repositories listed
-
Showing Your Offline Reinforcement Learning Work: Online Evaluation Budget Matters8 Oct 2021 0 repositories listed
-
Offline RL With Resource Constrained Online Deployment7 Oct 2021 0 repositories listed
-
You Only Evaluate Once: a Simple Baseline Algorithm for Offline RL5 Oct 2021 0 repositories listed
-
Adaptive Q-learning for Interaction-Limited Reinforcement Learning29 Sep 2021 0 repositories listed
-
CrowdPlay: Crowdsourcing human demonstration data for offline learning in Atari games29 Sep 2021 0 repositories listed
-
Data Sharing without Rewards in Multi-Task Offline Reinforcement Learning29 Sep 2021 0 repositories listed
-
Learning Pseudometric-based Action Representations for Offline Reinforcement Learning29 Sep 2021 0 repositories listed
-
Offline Reinforcement Learning for Large Scale Language Action Spaces29 Sep 2021 0 repositories listed
-
Offline Reinforcement Learning with Resource Constrained Online Deployment29 Sep 2021 0 repositories listed
-
Pareto Policy Pool for Model-based Offline Reinforcement Learning29 Sep 2021 0 repositories listed
-
29 Sep 2021 0 repositories listed
-
Reward Shifting for Optimistic Exploration and Conservative Exploitation29 Sep 2021 0 repositories listed
-
Semi-supervised Offline Reinforcement Learning with Pre-trained Decision Transformers29 Sep 2021 0 repositories listed
-
Should I Run Offline Reinforcement Learning or Behavioral Cloning?29 Sep 2021 0 repositories listed
-
Targeted Environment Design from Offline Data29 Sep 2021 0 repositories listed
-
The Essential Elements of Offline RL via Supervised Learning29 Sep 2021 0 repositories listed
-
Uncertainty Regularized Policy Learning for Offline Reinforcement Learning29 Sep 2021 0 repositories listed
-
Variational oracle guiding for reinforcement learning29 Sep 2021 0 repositories listed
Syntology lines on 3 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.