Browse State-of-the-Art › Sequential Decision Making › Papers, page 11
Sequential Decision Making
Papers archive 2025-07-28
archive papers tagged: 1,210 · with a code link: 351 · where Syntology ran a sample: 107 (90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (107 of 1,210 tagged: 90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 11 of 13: papers 1,001 to 1,100 of 1,210, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Circuit Routing Using Monte Carlo Tree Search and Deep Neural Networks24 Jun 2020 0 repositories listed
-
Risk-Sensitive Reinforcement Learning: a Martingale Approach to Reward Uncertainty23 Jun 2020 0 repositories listed
-
Towards Tractable Optimism in Model-Based Reinforcement Learning21 Jun 2020 0 repositories listed
-
Counterfactually Guided Off-policy Transfer in Clinical Settings20 Jun 2020 0 repositories listed
-
Learning by Repetition: Stochastic Multi-armed Bandits under Priming Effect18 Jun 2020 0 repositories listed
-
Parameterized MDPs and Reinforcement Learning Problems -- A Maximum Entropy Principle Based Framework17 Jun 2020 0 repositories listed
-
On the Relationship Between Structure in Natural Language and Models of Sequential Decision Processes12 Jun 2020 0 repositories listed
-
Group-Fair Online Allocation in Continuous Time11 Jun 2020 0 repositories listed
-
Modeling Human Driving Behavior through Generative Adversarial Imitation Learning10 Jun 2020 0 repositories listed
-
When is Particle Filtering Efficient for Planning in Partially Observed Linear Dynamical Systems?10 Jun 2020 0 repositories listed
-
Stealing Deep Reinforcement Learning Models for Fun and Profit9 Jun 2020 0 repositories listed
-
Sharp Thresholds of the Information Cascade Fragility Under a Mismatched Model7 Jun 2020 0 repositories listed
-
When Does MAML Objective Have Benign Landscape?31 May 2020 0 repositories listed
-
Dynamic Bi-Objective Routing of Multiple Vehicles28 May 2020 0 repositories listed
-
Active Measure Reinforcement Learning for Observation Cost Minimization26 May 2020 0 repositories listed
-
Causal Bayesian Optimization24 May 2020 0 repositories listed
-
Implementability of Honest Multi-Agent Sequential Decision-Making with Dynamic Population19 May 2020 0 repositories listed
-
Scalable First-Order Methods for Robust MDPs11 May 2020 0 repositories listed
-
Divide-and-Conquer Monte Carlo Tree Search For Goal-Directed Planning23 Apr 2020 0 repositories listed
-
iCORPP: Interleaved Commonsense Reasoning and Probabilistic Planning on Robots18 Apr 2020 0 repositories listed
-
Actor-Critic Deep Reinforcement Learning for Solving Job Shop Scheduling Problems14 Apr 2020 0 repositories listed
-
Sequential Batch Learning in Finite-Action Linear Contextual Bandits14 Apr 2020 0 repositories listed
-
A Deep Reinforcement Learning Framework for Continuous Intraday Market Bidding13 Apr 2020 0 repositories listed
-
Distributed Learning: Sequential Decision Making in Resource-Constrained Environments13 Apr 2020 0 repositories listed
-
Human AI interaction loop training: New approach for interactive reinforcement learning9 Mar 2020 0 repositories listed
-
A Farewell to Arms: Sequential Reward Maximization on a Budget with a Giving Up Option6 Mar 2020 0 repositories listed
-
Distributional Robustness and Regularization in Reinforcement Learning5 Mar 2020 0 repositories listed
-
Exploration-Exploitation in Constrained MDPs4 Mar 2020 0 repositories listed
-
Structure-Adaptive Sequential Testing for Online False Discovery Rate Control28 Feb 2020 0 repositories listed
-
Information Directed Sampling for Linear Partial Monitoring25 Feb 2020 0 repositories listed
-
Online Batch Decision-Making with High-Dimensional Covariates21 Feb 2020 0 repositories listed
-
Weakly-supervised Multi-output Regression via Correlated Gaussian Processes19 Feb 2020 0 repositories listed
-
Legion: Best-First Concolic Testing15 Feb 2020 0 repositories listed
-
Improving Generalization of Reinforcement Learning with Minimax Distributional Soft Actor-Critic13 Feb 2020 0 repositories listed
-
Listwise Learning to Rank with Deep Q-Networks13 Feb 2020 0 repositories listed
-
Tight Lower Bounds for Combinatorial Multi-Armed Bandits13 Feb 2020 0 repositories listed
-
Verifiable RNN-Based Policies for POMDPs Under Temporal Logic Constraints13 Feb 2020 0 repositories listed
-
Accelerating Reinforcement Learning for Reaching using Continuous Curriculum Learning7 Feb 2020 0 repositories listed
-
Bridging the Gap: Providing Post-Hoc Symbolic Explanations for Sequential Decision-Making Problems with Inscrutable Representations4 Feb 2020 0 repositories listed
-
Fairness in Learning-Based Sequential Decision Algorithms: A Survey14 Jan 2020 0 repositories listed
-
A storage expansion planning framework using reinforcement learning and simulation-based optimization10 Jan 2020 0 repositories listed
-
On Computation and Generalization of Generative Adversarial Imitation Learning9 Jan 2020 0 repositories listed
-
Direct and indirect reinforcement learning23 Dec 2019 0 repositories listed
-
Decentralized Multi-Agent Reinforcement Learning with Networked Agents: Recent Advances9 Dec 2019 0 repositories listed
-
Maximum Entropy Monte-Carlo Planning1 Dec 2019 0 repositories listed
-
An Optimized and Energy-Efficient Parallel Implementation of Non-Iteratively Trained Recurrent Neural Networks26 Nov 2019 0 repositories listed
-
Multi-Agent Reinforcement Learning: A Selective Overview of Theories and Algorithms24 Nov 2019 0 repositories listed
-
Working Memory Graphs17 Nov 2019 0 repositories listed
-
One-shot learning and behavioral eligibility traces in sequential decision making12 Nov 2019 0 repositories listed
-
Adaptivity in Adaptive Submodularity9 Nov 2019 0 repositories listed
-
Beta DVBF: Learning State-Space Models for Control from High Dimensional Observations2 Nov 2019 0 repositories listed
-
Adaptive Exploration in Linear Contextual Bandit15 Oct 2019 0 repositories listed
-
Reinforcement Learning for Multi-Objective Optimization of Online Decisions in High-Dimensional Systems1 Oct 2019 0 repositories listed
-
The Choice Function Framework for Online Policy Improvement1 Oct 2019 0 repositories listed
-
Collaborative Inter-agent Knowledge Distillation for Reinforcement Learning25 Sep 2019 0 repositories listed
-
Generalizing Reinforcement Learning to Unseen Actions25 Sep 2019 0 repositories listed
-
Learning Functionally Decomposed Hierarchies for Continuous Navigation Tasks25 Sep 2019 0 repositories listed
-
PROVABLY BENEFITS OF DEEP HIERARCHICAL RL25 Sep 2019 0 repositories listed
-
Selective Network Discovery via Deep Reinforcement Learning on Embedded Spaces16 Sep 2019 0 repositories listed
-
An Arm-Wise Randomization Approach to Combinatorial Linear Semi-Bandits5 Sep 2019 0 repositories listed
-
Can A User Anticipate What Her Followers Want?1 Sep 2019 0 repositories listed
-
Reinforcement Learning in Healthcare: A Survey22 Aug 2019 0 repositories listed
-
Exploring Offline Policy Evaluation for the Continuous-Armed Bandit Problem21 Aug 2019 0 repositories listed
-
Online Planning for Decentralized Stochastic Control with Partial History Sharing6 Aug 2019 0 repositories listed
-
Bridging Commonsense Reasoning and Probabilistic Planning via a Probabilistic Action Language31 Jul 2019 0 repositories listed
-
Bandit Convex Optimization in Non-stationary Environments29 Jul 2019 0 repositories listed
-
IR-VIC: Unsupervised Discovery of Sub-goals for Transfer in RL24 Jul 2019 0 repositories listed
-
A Sufficient Statistic for Influence in Structured Multiagent Environments22 Jul 2019 0 repositories listed
-
Reward Advancement: Transforming Policy under Maximum Causal Entropy Principle11 Jul 2019 0 repositories listed
-
A Scheme for Dynamic Risk-Sensitive Sequential Decision Making9 Jul 2019 0 repositories listed
-
Thompson Sampling on Symmetric α-Stable Bandits8 Jul 2019 0 repositories listed
-
Exploiting Relevance for Online Decision-Making in High-Dimensions1 Jul 2019 0 repositories listed
-
Learning Markov models via low-rank optimization28 Jun 2019 0 repositories listed
-
A Theoretical Connection Between Statistical Physics and Reinforcement Learning24 Jun 2019 0 repositories listed
-
Macro-action Multi-time scale Dynamic Programming for Energy Management in Buildings with Phase Change Materials11 Jun 2019 0 repositories listed
-
Neural Heterogeneous Scheduler9 Jun 2019 0 repositories listed
-
Non-Stationary Reinforcement Learning: The Blessing of (More) Optimism7 Jun 2019 0 repositories listed
-
Learning NP-Hard Multi-Agent Assignment Planning using GNN: Inference on a Random Graph and Provable Auction-Fitted Q-learning29 May 2019 0 repositories listed
-
Knowledge-Based Sequential Decision-Making Under Uncertainty16 May 2019 0 repositories listed
-
Tight Regret Bounds for Infinite-armed Linear Contextual Bandits4 May 2019 0 repositories listed
-
Group Retention when Using Machine Learning in Sequential Decision Making: the Interplay between User Dynamics and Fairness2 May 2019 0 repositories listed
-
Soft Q-Learning with Mutual-Information Regularization1 May 2019 0 repositories listed
-
Trajectory VAE for multi-modal imitation1 May 2019 0 repositories listed
-
Understanding & Generalizing AlphaGo Zero1 May 2019 0 repositories listed
-
Deep Reinforcement Learning for Optimal Critical Care Pain Management with Morphine using Dueling Double-Deep Q Networks25 Apr 2019 0 repositories listed
-
Beyond Adaptive Submodularity: Approximation Guarantees of Greedy Policy with Adaptive Submodularity Ratio24 Apr 2019 0 repositories listed
-
Latent Variable Algorithms for Multimodal Learning and Sensor Fusion23 Apr 2019 0 repositories listed
-
A Short Survey On Memory Based Reinforcement Learning14 Apr 2019 0 repositories listed
-
Similarities between policy gradient methods (PGM) in Reinforcement learning (RL) and supervised learning (SL)12 Apr 2019 0 repositories listed
-
Meta-Learning surrogate models for sequential decision making28 Mar 2019 0 repositories listed
-
Automating Predictive Modeling Process using Reinforcement Learning2 Mar 2019 0 repositories listed
-
Design of intentional backdoors in sequential models26 Feb 2019 0 repositories listed
-
Human-in-the-loop Active Covariance Learning for Improving Prediction in Small Data Sets26 Feb 2019 0 repositories listed
-
Scalable Thompson Sampling via Optimal Transport19 Feb 2019 0 repositories listed
-
Network Offloading Policies for Cloud Robotics: a Learning-based Approach15 Feb 2019 0 repositories listed
-
Rethinking the Discount Factor in Reinforcement Learning: A Decision Theoretic Approach8 Feb 2019 0 repositories listed
-
Hyper-parameter Tuning under a Budget Constraint1 Feb 2019 0 repositories listed
-
Deep Neural Linear Bandits: Overcoming Catastrophic Forgetting through Likelihood Matching24 Jan 2019 0 repositories listed
-
Learning and Reasoning for Robot Sequential Decision Making under Uncertainty16 Jan 2019 0 repositories listed
-
A Computational Framework for Motor Skill Acquisition3 Jan 2019 0 repositories listed