Browse State-of-the-Art › Sequential Decision Making › Papers, page 6
Sequential Decision Making
Papers archive 2025-07-28
archive papers tagged: 1,210 · with a code link: 351 · where Syntology ran a sample: 107 (90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (107 of 1,210 tagged: 90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 6 of 13: papers 501 to 600 of 1,210, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
How to Measure Human-AI Prediction Accuracy in Explainable AI Systems23 Aug 2024 0 repositories listed
-
Pareto Inverse Reinforcement Learning for Diverse Expert Policy Generation22 Aug 2024 0 repositories listed
-
An End-to-End Reinforcement Learning Based Approach for Micro-View Order-Dispatching in Ride-Hailing20 Aug 2024 0 repositories listed
-
Contextual Bandits for Unbounded Context Distributions19 Aug 2024 0 repositories listed
-
Meta Clustering of Neural Bandits10 Aug 2024 0 repositories listed
-
Structure and Reduction of MCTS for Explainable-AI10 Aug 2024 0 repositories listed
-
Non-maximizing policies that fulfill multi-criterion aspirations in expectation8 Aug 2024 0 repositories listed
-
Few-shot Scooping Under Domain Shift via Simulated Maximal Deployment Gaps6 Aug 2024 0 repositories listed
-
How to Choose a Reinforcement-Learning Algorithm30 Jul 2024 0 repositories listed
-
Reinforcement Learning for Sustainable Energy: A Survey26 Jul 2024 0 repositories listed
-
Assessing AI Utility: The Random Guesser Test for Sequential Decision-Making Systems25 Jul 2024 0 repositories listed
-
Differentiable Quantum Architecture Search in Asynchronous Quantum Reinforcement Learning25 Jul 2024 0 repositories listed
-
Is Behavior Cloning All You Need? Understanding Horizon in Imitation Learning20 Jul 2024 0 repositories listed
-
Managing Risk using Rolling Forecasts in Energy-Limited and Stochastic Energy Systems18 Jul 2024 0 repositories listed
-
Exploration Unbound16 Jul 2024 0 repositories listed
-
Interpretability in Action: Exploratory Analysis of VPT, a Minecraft Agent16 Jul 2024 0 repositories listed
-
Predicting and Understanding Human Action Decisions: Insights from Large Language Models and Cognitive Instance-Based Learning12 Jul 2024 0 repositories listed
-
MDP Geometry, Normalization and Reward Balancing Solvers9 Jul 2024 0 repositories listed
-
Communication and Control Co-Design in 6G: Sequential Decision-Making with LLMs6 Jul 2024 0 repositories listed
-
Maximizing utility in multi-agent environments by anticipating the behavior of other learners5 Jul 2024 0 repositories listed
-
Short-Long Policy Evaluation with Novel Actions4 Jul 2024 0 repositories listed
-
Exploring a Physics-Informed Decision Transformer for Distribution System Restoration: Methodology and Performance Analysis30 Jun 2024 0 repositories listed
-
Tradeoffs When Considering Deep Reinforcement Learning for Contingency Management in Advanced Air Mobility28 Jun 2024 0 repositories listed
-
UNO Arena for Evaluating Sequential Decision-Making Capability of Large Language Models24 Jun 2024 0 repositories listed
-
Accelerating Matrix Diagonalization through Decision Transformers with Epsilon-Greedy Optimization23 Jun 2024 0 repositories listed
-
ARDuP: Active Region Video Diffusion for Universal Policies19 Jun 2024 0 repositories listed
-
Learned Graph Rewriting with Equality Saturation: A New Paradigm in Relational Query Rewrite and Beyond19 Jun 2024 0 repositories listed
-
Constrained Reinforcement Learning with Average Reward Objective: Model-Based and Model-Free Algorithms17 Jun 2024 0 repositories listed
-
Efficient Sequential Decision Making with Large Language Models17 Jun 2024 0 repositories listed
-
Model Adaptation for Time Constrained Embodied Control17 Jun 2024 0 repositories listed
-
Data-Driven Upper Confidence Bounds with Near-Optimal Regret for Heavy-Tailed Bandits9 Jun 2024 0 repositories listed
-
Rectifying Reinforcement Learning for Reward Matching4 Jun 2024 0 repositories listed
-
Low-rank finetuning for LLMs: A fairness perspective28 May 2024 0 repositories listed
-
Leveraging Offline Data in Linear Latent Bandits27 May 2024 0 repositories listed
-
OPERA: Automatic Offline Policy Evaluation with Re-weighted Aggregates of Multiple Estimators27 May 2024 0 repositories listed
-
Variational Offline Multi-agent Skill Discovery26 May 2024 0 repositories listed
-
Inference of Utilities and Time Preference in Sequential Decision-Making24 May 2024 0 repositories listed
-
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning24 May 2024 0 repositories listed
-
A finite time analysis of distributed Q-learning23 May 2024 0 repositories listed
-
Reinforcing Language Agents via Policy Optimization with Action Decomposition23 May 2024 0 repositories listed
-
Efficiently Training Deep-Learning Parametric Policies using Lagrangian Duality23 May 2024 0 repositories listed
-
Understanding the Training and Generalization of Pretrained Transformer for Sequential Decision Making23 May 2024 0 repositories listed
-
On the Brittle Foundations of ReAct Prompting for Agentic Large Language Models22 May 2024 0 repositories listed
-
A Unified Linear Programming Framework for Offline Reward Learning from Human Demonstrations and Feedback20 May 2024 0 repositories listed
-
CPS-LLM: Large Language Model based Safe Usage Plan Generator for Human-in-the-Loop Human-in-the-Plant Cyber-Physical System19 May 2024 0 repositories listed
-
AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments13 May 2024 0 repositories listed
-
Human-Modeling in Sequential Decision-Making: An Analysis through the Lens of Human-Aware AI13 May 2024 0 repositories listed
-
Enhancing Q-Learning with Large Language Model Heuristics6 May 2024 0 repositories listed
-
Learning Planning Abstractions from Language6 May 2024 0 repositories listed
-
Out-of-Distribution Adaptation in Offline RL: Counterfactual Reasoning via Causal Normalizing Flows6 May 2024 0 repositories listed
-
MEXGEN: An Effective and Efficient Information Gain Approximation for Information Gathering Path Planning4 May 2024 0 repositories listed
-
Mathematics of statistical sequential decision-making: concentration, risk-awareness and modelling in stochastic bandits, with applications to bariatric surgery3 May 2024 0 repositories listed
-
Provably Efficient Reinforcement Learning for Adversarial Restless Multi-Armed Bandits with Unknown Transitions and Bandit Feedback2 May 2024 0 repositories listed
-
Scalable Bayesian Inference in the Era of Deep Learning: From Gaussian Processes to Deep Neural Networks29 Apr 2024 0 repositories listed
-
Q-learning with temporal memory to navigate turbulence26 Apr 2024 0 repositories listed
-
Digital Twins for forecasting and decision optimisation with machine learning: applications in wastewater treatment23 Apr 2024 0 repositories listed
-
Do LLMs Play Dice? Exploring Probability Distribution Sampling in Large Language Models for Behavioral Simulation13 Apr 2024 0 repositories listed
-
Reward Learning from Suboptimal Demonstrations with Applications in Surgical Electrocautery10 Apr 2024 0 repositories listed
-
Regularized Conditional Diffusion Model for Multi-Task Preference Alignment7 Apr 2024 0 repositories listed
-
Composite Bayesian Optimization In Function Spaces Using NEON -- Neural Epistemic Operator Networks3 Apr 2024 0 repositories listed
-
Multi-granular Adversarial Attacks against Black-box Neural Ranking Models2 Apr 2024 0 repositories listed
-
Retentive Decision Transformer with Adaptive Masking for Reinforcement Learning based Recommendation Systems26 Mar 2024 0 repositories listed
-
Continual Vision-and-Language Navigation22 Mar 2024 0 repositories listed
-
Sequential Decision-Making for Inline Text Autocomplete21 Mar 2024 0 repositories listed
-
Fast Value Tracking for Deep Reinforcement Learning19 Mar 2024 0 repositories listed
-
State-Separated SARSA: A Practical Sequential Decision-Making Algorithm with Recovering Rewards18 Mar 2024 0 repositories listed
-
Supervised Fine-Tuning as Inverse Reinforcement Learning18 Mar 2024 0 repositories listed
-
Distributed Multi-Objective Dynamic Offloading Scheduling for Air-Ground Cooperative MEC16 Mar 2024 0 repositories listed
-
Regret Minimization via Saddle Point Optimization15 Mar 2024 0 repositories listed
-
AutoGuide: Automated Generation and Selection of Context-Aware Guidelines for Large Language Model Agents13 Mar 2024 0 repositories listed
-
CoRAL: Collaborative Retrieval-Augmented Large Language Models Improve Long-tail Recommendation11 Mar 2024 0 repositories listed
-
LinearAPT: An Adaptive Algorithm for the Fixed-Budget Thresholding Linear Bandit Problem10 Mar 2024 0 repositories listed
-
Inverse Design of Photonic Crystal Surface Emitting Lasers is a Sequence Modeling Problem8 Mar 2024 0 repositories listed
-
Cooperative Bayesian Optimization for Imperfect Agents7 Mar 2024 0 repositories listed
-
A Survey on Applications of Reinforcement Learning in Spatial Resource Allocation6 Mar 2024 0 repositories listed
-
Language Guided Exploration for RL Agents in Text Environments5 Mar 2024 0 repositories listed
-
Adaptive Learning Rate for Follow-the-Regularized-Leader: Competitive Analysis and Best-of-Both-Worlds1 Mar 2024 0 repositories listed
-
On the Role of Information Structure in Reinforcement Learning for Partially-Observable Sequential Teams and Games1 Mar 2024 0 repositories listed
-
Successfully Guiding Humans with Imperfect Instructions by Highlighting Potential Errors and Suggesting Corrections26 Feb 2024 0 repositories listed
-
Information-Theoretic Safe Bayesian Optimization23 Feb 2024 0 repositories listed
-
BeTAIL: Behavior Transformer Adversarial Imitation Learning from Human Racing Gameplay22 Feb 2024 0 repositories listed
-
On the Performance of Empirical Risk Minimization with Smoothed Data22 Feb 2024 0 repositories listed
-
Align Your Intents: Offline Imitation Learning via Optimal Transport20 Feb 2024 0 repositories listed
-
Toward TransfORmers: Revolutionizing the Solution of Mixed Integer Programs with Transformers20 Feb 2024 0 repositories listed
-
Self-evolving Autoencoder Embedded Q-Network18 Feb 2024 0 repositories listed
-
Probability Tools for Sequential Random Projection16 Feb 2024 0 repositories listed
-
Auxiliary Reward Generation with Transition Distance Representation Learning12 Feb 2024 0 repositories listed
-
Online Sequential Decision-Making with Unknown Delays12 Feb 2024 0 repositories listed
-
Offline Risk-sensitive RL with Partial Observability to Enhance Performance in Human-Robot Teaming8 Feb 2024 0 repositories listed
-
A Reinforcement Learning Approach for Dynamic Rebalancing in Bike-Sharing System5 Feb 2024 0 repositories listed
-
Multi-Agent Reinforcement Learning for Offloading Cellular Communications with Cooperating UAVs5 Feb 2024 0 repositories listed
-
Regularized Q-Learning with Linear Function Approximation26 Jan 2024 0 repositories listed
-
Stochastic Dynamic Power Dispatch with High Generalization and Few-Shot Adaption via Contextual Meta Graph Reinforcement Learning19 Jan 2024 0 repositories listed
-
LLMs for Relational Reasoning: How Far are We?17 Jan 2024 0 repositories listed
-
Towards Off-Policy Reinforcement Learning for Ranking Policies with Human Feedback17 Jan 2024 0 repositories listed
-
Graph Q-Learning for Combinatorial Optimization11 Jan 2024 0 repositories listed
-
Interactions between dynamic team composition and coordination: An agent-based modeling approach11 Jan 2024 0 repositories listed
-
On Sample-Efficient Offline Reinforcement Learning: Data Diversity, Posterior Sampling, and Beyond6 Jan 2024 0 repositories listed
-
Towards an Adaptable and Generalizable Optimization Engine in Decision and Control: A Meta Reinforcement Learning Approach4 Jan 2024 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.