Browse State-of-the-Art › Model-based Reinforcement Learning › Papers, page 4
Model-based Reinforcement Learning
Papers archive 2025-07-28
archive papers tagged: 708 · with a code link: 234 · where Syntology ran a sample: 72 (66 with a run with no instrument failure, 6 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (72 of 708 tagged: 66 with a run with no instrument failure, 6 where every run was a failure of Syntology's instrument)
Page 4 of 8: papers 301 to 400 of 708, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Model-based Policy Optimization using Symbolic World Model18 Jul 2024 0 repositories listed
-
BECAUSE: Bilinear Causal Representation for Generalizable Offline Model-based Reinforcement Learning15 Jul 2024 0 repositories listed
-
Graph Neural Networks with Model-based Reinforcement Learning for Multi-agent Systems12 Jul 2024 0 repositories listed
-
FOSP: Fine-tuning Offline Safe Policy through World Models6 Jul 2024 0 repositories listed
-
Hybrid RAG-empowered Multi-modal LLM for Secure Data Management in Internet of Medical Things: A Diffusion-based Contract Approach1 Jul 2024 0 repositories listed
-
Identifying Ordinary Differential Equations for Data-efficient Model-based Reinforcement Learning28 Jun 2024 0 repositories listed
-
State and Input Constrained Output-Feedback Adaptive Optimal Control of Affine Nonlinear Systems27 Jun 2024 0 repositories listed
-
Infinite-Horizon Reinforcement Learning with Multinomial Logistic Function Approximation19 Jun 2024 0 repositories listed
-
Reinforcement learning-based architecture search for quantum machine learning4 Jun 2024 0 repositories listed
-
A New View on Planning in Online Reinforcement Learning3 Jun 2024 0 repositories listed
-
Adaptive Layer Splitting for Wireless LLM Inference in Edge Computing: A Model-Based Reinforcement Learning Approach3 Jun 2024 0 repositories listed
-
Learning to Play Atari in a World of Tokens3 Jun 2024 0 repositories listed
-
Adaptive Horizon Actor-Critic for Policy Learning in Contact-Rich Differentiable Simulation28 May 2024 0 repositories listed
-
Partial Models for Building Adaptive Model-Based Reinforcement Learning Agents27 May 2024 0 repositories listed
-
Model-based reinforcement learning for protein backbone design3 May 2024 0 repositories listed
-
Continual Model-based Reinforcement Learning for Data Efficient Wireless Network Optimisation30 Apr 2024 0 repositories listed
-
A Note on Loss Functions and Error Compounding in Model-based Reinforcement Learning15 Apr 2024 0 repositories listed
-
Active Learning for Control-Oriented Identification of Nonlinear Systems13 Apr 2024 0 repositories listed
-
Active Exploration in Bayesian Model-based Reinforcement Learning for Robot Manipulation2 Apr 2024 0 repositories listed
-
Robust Model Based Reinforcement Learning Using ℒ₁ Adaptive Control21 Mar 2024 0 repositories listed
-
Dr. Strategy: Model-Based Generalist Agents with Strategic Dreaming29 Feb 2024 0 repositories listed
-
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination23 Feb 2024 0 repositories listed
-
Model-Based Reinforcement Learning Control of Reaction-Diffusion Problems22 Feb 2024 0 repositories listed
-
Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption14 Feb 2024 0 repositories listed
-
Differentially Private Deep Model-Based Reinforcement Learning8 Feb 2024 0 repositories listed
-
A Multi-step Loss Function for Robust Learning of the Dynamics in Model-based Reinforcement Learning5 Feb 2024 0 repositories listed
-
Deep autoregressive density nets vs neural ensembles for model-based offline reinforcement learning5 Feb 2024 0 repositories listed
-
Control in Stochastic Environment with Delays: A Model-based Reinforcement Learning Approach1 Feb 2024 0 repositories listed
-
Scheduled Curiosity-Deep Dyna-Q: Efficient Exploration for Dialog Policy Learning31 Jan 2024 0 repositories listed
-
Locality Sensitive Sparse Encoding for Learning World Models Online23 Jan 2024 0 repositories listed
-
Emergence and Causality in Complex Systems: A Survey on Causal Emergence and Related Quantitative Studies28 Dec 2023 0 repositories listed
-
Dynamic Programming-based Approximate Optimal Control for Model-Based Reinforcement Learning22 Dec 2023 0 repositories listed
-
Model-Based Epistemic Variance of Values for Risk-Aware Policy Optimization7 Dec 2023 0 repositories listed
-
Score-Aware Policy-Gradient Methods and Performance Guarantees using Local Lyapunov Conditions: Applications to Product-Form Stochastic Networks and Queueing Systems5 Dec 2023 0 repositories listed
-
3 Dec 2023 0 repositories listed Syntology 9 ran (of which 0 constructed an object rather than computing a result; 9 with no instrument failure: 0 honoured, 0 violated, 9 with no contract checked; 0 where Syntology's instrument failed) · 7 unverified (of 16 harvested samples) · 16 pointer-only (licence)
-
LanGWM: Language Grounded World Model29 Nov 2023 0 repositories listed
-
Autonomous Port Navigation With Ranging Sensors Using Model-Based Reinforcement Learning17 Nov 2023 0 repositories listed
-
An introduction to reinforcement learning for neuroscience13 Nov 2023 0 repositories listed
-
Reinforcement Twinning: from digital twins to model-based reinforcement learning7 Nov 2023 0 repositories listed
-
DreamSmooth: Improving Model-based Reinforcement Learning via Reward Smoothing2 Nov 2023 0 repositories listed
-
The Alignment Ceiling: Objective Mismatch in Reinforcement Learning from Human Feedback31 Oct 2023 0 repositories listed
-
Efficient Exploration in Continuous-time Model-based Reinforcement Learning30 Oct 2023 0 repositories listed
-
Relational Object-Centric Actor-Critic26 Oct 2023 0 repositories listed
-
Mind the Model, Not the Agent: The Primacy Bias in Model-based RL23 Oct 2023 0 repositories listed
-
Tree Search in DAG Space with Model-based Reinforcement Learning for Causal Discovery20 Oct 2023 0 repositories listed
-
Value-Biased Maximum Likelihood Estimation for Model-based Reinforcement Learning in Discounted Linear MDPs17 Oct 2023 0 repositories listed
-
MoConVQ: Unified Physics-Based Motion Control via Scalable Discrete Representations16 Oct 2023 0 repositories listed
-
COPlanner: Plan to Roll Out Conservatively but to Explore Optimistically for Model-Based RL11 Oct 2023 0 repositories listed
-
Multi-timestep models for Model-based Reinforcement Learning9 Oct 2023 0 repositories listed
-
Reward-Consistent Dynamics Models are Strongly Generalizable for Offline Reinforcement Learning9 Oct 2023 0 repositories listed
-
Amortized Network Intervention to Steer the Excitatory Point Processes6 Oct 2023 0 repositories listed
-
MoDem-V2: Visuo-Motor World Models for Real-World Robot Manipulation25 Sep 2023 0 repositories listed
-
DOMAIN: MilDly COnservative Model-BAsed OfflINe Reinforcement Learning16 Sep 2023 0 repositories listed
-
Mind the Uncertainty: Risk-Aware and Actively Exploring Model-Based Reinforcement Learning11 Sep 2023 0 repositories listed
-
Distributionally Robust Model-based Reinforcement Learning with Large State Spaces5 Sep 2023 0 repositories listed
-
Exploring the Potential of World Models for Anomaly Detection in Autonomous Driving10 Aug 2023 0 repositories listed
-
Theoretically Guaranteed Policy Improvement Distilled from Model-Based Planning24 Jul 2023 0 repositories listed
-
Image Transformation Sequence Retrieval with General Reinforcement Learning13 Jul 2023 0 repositories listed
-
Facing Off World Model Backbones: RNNs, Transformers, and S45 Jul 2023 0 repositories listed
-
Surge Routing: Event-informed Multiagent Reinforcement Learning for Autonomous Rideshare5 Jul 2023 0 repositories listed
-
λ-models: Effective Decision-Aware Reinforcement Learning with Latent Models30 Jun 2023 0 repositories listed
-
Deep Generative Models for Decision-Making and Control15 Jun 2023 0 repositories listed
-
How to Learn and Generalize From Three Minutes of Data: Physics-Constrained and Uncertainty-Aware Neural Stochastic Differential Equations10 Jun 2023 0 repositories listed
-
IQL-TD-MPC: Implicit Q-Learning for Hierarchical Model Predictive Control1 Jun 2023 0 repositories listed
-
What model does MuZero learn?1 Jun 2023 0 repositories listed
-
Digital Twin-Based 3D Map Management for Edge-Assisted Mobile Augmented Reality26 May 2023 0 repositories listed
-
TOM: Learning Policy-Aware Models for Model-Based Reinforcement Learning via Transition Occupancy Matching22 May 2023 0 repositories listed
-
Bridging Active Exploration and Uncertainty-Aware Deployment Using Probabilistic Ensemble Neural Network Dynamics20 May 2023 0 repositories listed
-
Sense, Imagine, Act: Multimodal Perception Improves Model-Based Reinforcement Learning for Head-to-Head Autonomous Racing8 May 2023 0 repositories listed
-
A Survey on Offline Model-Based Reinforcement Learning5 May 2023 0 repositories listed
-
Human Machine Co-adaption Interface via Cooperation Markov Decision Process System3 May 2023 0 repositories listed
-
Model Based Reinforcement Learning for Personalized Heparin Dosing19 Apr 2023 0 repositories listed
-
Decision-Focused Model-based Reinforcement Learning for Reward Transfer6 Apr 2023 0 repositories listed
-
State and Parameter Estimation for Affine Nonlinear Systems4 Apr 2023 0 repositories listed
-
Risk-Sensitive and Robust Model-Based Reinforcement Learning and Planning2 Apr 2023 0 repositories listed
-
EDGI: Equivariant Diffusion for Planning with Embodied Agents22 Mar 2023 0 repositories listed
-
A New Policy Iteration Algorithm For Reinforcement Learning in Zero-Sum Markov Games17 Mar 2023 0 repositories listed
-
On the Benefits of Leveraging Structural Information in Planning Over the Learned Model15 Mar 2023 0 repositories listed
-
Replay Buffer with Local Forgetting for Adapting to Local Environment Changes in Deep Model-Based Reinforcement Learning15 Mar 2023 0 repositories listed
-
Beware of Instantaneous Dependence in Reinforcement Learning9 Mar 2023 0 repositories listed
-
Approximating Energy Market Clearing and Bidding With Model-Based Reinforcement Learning3 Mar 2023 0 repositories listed
-
Understanding the effect of varying amounts of replay per step20 Feb 2023 0 repositories listed
-
Equivariant MuZero9 Feb 2023 0 repositories listed
-
Is Model Ensemble Necessary? Model-based RL via a Single Model with Lipschitz Regularized Value Function2 Feb 2023 0 repositories listed
-
Learning Control from Raw Position Measurements30 Jan 2023 0 repositories listed
-
STEERING: Stein Information Directed Exploration for Model-Based Reinforcement Learning28 Jan 2023 0 repositories listed
-
Intrinsic Motivation in Model-based Reinforcement Learning: A Brief Review24 Jan 2023 0 repositories listed
-
Minimal Value-Equivalent Partial Models for Scalable and Robust Planning in Lifelong Reinforcement Learning24 Jan 2023 0 repositories listed
-
Model Based Reinforcement Learning with Non-Gaussian Environment Dynamics and its Application to Portfolio Optimization23 Jan 2023 0 repositories listed
-
Exploration in Model-based Reinforcement Learning with Randomized Reward9 Jan 2023 0 repositories listed
-
Model-Based Reinforcement Learning with Multinomial Logistic Function Approximation27 Dec 2022 0 repositories listed
-
Latent Variable Representation for Reinforcement Learning17 Dec 2022 0 repositories listed
-
A Data-driven Pricing Scheme for Optimal Routing through Artificial Currencies27 Nov 2022 0 repositories listed
-
Prototypical context-aware dynamics generalization for high-dimensional model-based reinforcement learning23 Nov 2022 0 repositories listed
-
Offline Estimation of Controlled Markov Chains: Minimaxity and Sample Complexity14 Nov 2022 0 repositories listed
-
Contrastive Value Learning: Implicit Models for Simple Offline RL3 Nov 2022 0 repositories listed
-
Model-based Reinforcement Learning with a Hamiltonian Canonical ODE Network2 Nov 2022 0 repositories listed
-
SAM-RL: Sensing-Aware Model-Based Reinforcement Learning via Differentiable Physics-Based Simulation and Rendering27 Oct 2022 0 repositories listed
-
Active Exploration for Robotic Manipulation23 Oct 2022 0 repositories listed
-
Epistemic Monte Carlo Tree Search21 Oct 2022 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.