Browse State-of-the-Art › Sequential Decision Making › Papers, page 12
Sequential Decision Making
Papers archive 2025-07-28
archive papers tagged: 1,210 · with a code link: 351 · where Syntology ran a sample: 107 (90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (107 of 1,210 tagged: 90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 12 of 13: papers 1,101 to 1,200 of 1,210, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Deep Reinforcement Learning for Multi-Agent Systems: A Review of Challenges, Solutions and Applications31 Dec 2018 0 repositories listed
-
Monte-Carlo Tree Search for Constrained POMDPs1 Dec 2018 0 repositories listed
-
Negotiable Reinforcement Learning for Pareto Optimal Sequential Decision-Making1 Dec 2018 0 repositories listed
-
Regret Bounds for Online Portfolio Selection with a Cardinality Constraint1 Dec 2018 0 repositories listed
-
Tight Bayesian Ambiguity Sets for Robust MDPs15 Nov 2018 0 repositories listed
-
Meta-Learning for Multi-objective Reinforcement Learning8 Nov 2018 0 repositories listed
-
On preserving non-discrimination when combining expert advice28 Oct 2018 0 repositories listed
-
Resilient Computing with Reinforcement Learning on a Dynamical System: Case Study in Sorting25 Sep 2018 0 repositories listed
-
Geometric Multi-Model Fitting by Deep Reinforcement Learning22 Sep 2018 0 repositories listed
-
Predicting Periodicity with Temporal Difference Learning20 Sep 2018 0 repositories listed
-
Online Convex Optimization for Sequential Decision Processes and Extensive-Form Games10 Sep 2018 0 repositories listed
-
Investigating Order Effects in Multidimensional Relevance Judgment using Query Logs14 Jul 2018 0 repositories listed
-
Video Summarisation by Classification with Deep Reinforcement Learning9 Jul 2018 0 repositories listed
-
Playing against Nature: causal discovery for decision making under uncertainty3 Jul 2018 0 repositories listed
-
Multi-Task Generative Adversarial Nets with Shared Memory for Cross-Domain Coordination Control1 Jul 2018 0 repositories listed
-
Stagewise Safe Bayesian Optimization with Gaussian Processes20 Jun 2018 0 repositories listed
-
TopRank: A practical algorithm for online stochastic ranking6 Jun 2018 0 repositories listed
-
Learning Self-Imitating Diverse Policies25 May 2018 0 repositories listed
-
Fast Online Exact Solutions for Deterministic MDPs with Sparse Rewards8 May 2018 0 repositories listed
-
On Improving Deep Reinforcement Learning for POMDPs17 Apr 2018 0 repositories listed
-
UCBoost: A Boosting Approach to Tame Complexity and Optimality for Stochastic Bandits16 Apr 2018 0 repositories listed
-
Policy Gradient With Value Function Approximation For Collective Multiagent Planning9 Apr 2018 0 repositories listed
-
Hindsight is Only 50/50: Unsuitability of MDP based Approximate POMDP Solvers for Multi-resolution Information Gathering7 Apr 2018 0 repositories listed
-
Accelerating E-Commerce Search Engine Ranking by Contextual Factor Selection14 Mar 2018 0 repositories listed
-
Hierarchical Imitation and Reinforcement Learning1 Mar 2018 0 repositories listed
-
Novel Approaches to Accelerating the Convergence Rate of Markov Decision Process for Search Result Diversification23 Feb 2018 0 repositories listed
-
An Anytime Algorithm for Task and Motion MDPs16 Feb 2018 0 repositories listed
-
MPC-Inspired Neural Network Policies for Sequential Decision Making15 Feb 2018 0 repositories listed
-
Understanding Human Behaviors in Crowds by Imitating the Decision-Making Process25 Jan 2018 0 repositories listed
-
Testing Optimality of Sequential Decision-Making4 Jan 2018 0 repositories listed
-
Multi-shot Pedestrian Re-identification via Sequential Decision Making19 Dec 2017 0 repositories listed
-
Q-LDA: Uncovering Latent Patterns in Text-based Sequential Decision Processes1 Dec 2017 0 repositories listed
-
Loss Functions for Multiset Prediction14 Nov 2017 0 repositories listed
-
Servant of Many Masters: Shifting priorities in Pareto-optimal sequential decision-making31 Oct 2017 0 repositories listed
-
How Should a Robot Assess Risk? Towards an Axiomatic Theory of Risk in Robotics30 Oct 2017 0 repositories listed
-
Hierarchical State Abstractions for Decision-Making Problems with Computational Constraints22 Oct 2017 0 repositories listed
-
Asymmetric Actor Critic for Image-Based Robot Learning18 Oct 2017 0 repositories listed
-
Optimal Learning for Sequential Decision Making for Expensive Cost Functions with Stochastic Binary Feedbacks13 Sep 2017 0 repositories listed
-
Safety-Aware Algorithms for Adversarial Contextual Bandit1 Aug 2017 0 repositories listed
-
Non-Stationary Bandits with Habituation and Recovery Dynamics26 Jul 2017 0 repositories listed
-
Learning for Multi-robot Cooperation in Partially Observable Stochastic Environments with Macro-actions24 Jul 2017 0 repositories listed
-
Correlational Dueling Bandits with Application to Clinical Treatment in Large Decision Spaces8 Jul 2017 0 repositories listed
-
Tableaux for Policy Synthesis for MDPs with PCTL* Constraints30 Jun 2017 0 repositories listed
-
The Theory is Predictive, but is it Complete? An Application to Human Perception of Randomness21 Jun 2017 0 repositories listed
-
Unlocking the Potential of Simulators: Design with RL in Mind8 Jun 2017 0 repositories listed
-
A method for the online construction of the set of states of a Markov Decision Process using Answer Set Programming5 Jun 2017 0 repositories listed
-
Boltzmann Exploration Done Right29 May 2017 0 repositories listed
-
Learning to Mix n-Step Returns: Generalizing lambda-Returns for Deep Reinforcement Learning21 May 2017 0 repositories listed
-
Answer Set Programming for Non-Stationary Markov Decision Processes3 May 2017 0 repositories listed
-
Using Reinforcement Learning for Demand Response of Domestic Hot Water Buffers: a Real-Life Demonstration16 Mar 2017 0 repositories listed
-
Minimizing Maximum Regret in Commitment Constrained Sequential Decision Making14 Mar 2017 0 repositories listed
-
Deep Robust Kalman Filter7 Mar 2017 0 repositories listed
-
Deeply AggreVaTeD: Differentiable Imitation Learning for Sequential Prediction3 Mar 2017 0 repositories listed
-
Active Learning for Accurate Estimation of Linear Models2 Mar 2017 0 repositories listed
-
Tight Bounds for Bandit Combinatorial Optimization24 Feb 2017 0 repositories listed
-
Learning to Repeat: Fine Grained Action Repetition for Deep Reinforcement Learning20 Feb 2017 0 repositories listed
-
Deep Reinforcement Learning for Visual Object Tracking in Videos31 Jan 2017 0 repositories listed
-
Model-Free Control of Thermostatically Controlled Loads Connected to a District Heating Network27 Jan 2017 0 repositories listed
-
A Contextual Bandit Approach for Stream-Based Active Learning24 Jan 2017 0 repositories listed
-
Toward negotiable reinforcement learning: shifting priorities in Pareto optimal sequential decision-making5 Jan 2017 0 repositories listed
-
Stochastic Planning and Lifted Inference4 Jan 2017 0 repositories listed
-
From Preference-Based to Multiobjective Sequential Decision-Making3 Jan 2017 0 repositories listed
-
Multi-armed Bandits: Competing with Optimal Sequences1 Dec 2016 0 repositories listed
-
Fast Video Classification via Adaptive Cascading of Deep Models20 Nov 2016 0 repositories listed
-
Open Problem: Approximate Planning of POMDPs in the class of Memoryless Policies17 Aug 2016 0 repositories listed
-
Human collective intelligence as distributed Bayesian inference5 Aug 2016 0 repositories listed
-
Safe Policy Improvement by Minimizing Robust Baseline Regret13 Jul 2016 0 repositories listed
-
Preference at First Sight24 Jun 2016 0 repositories listed
-
The Bayesian Linear Information Filtering Problem30 May 2016 0 repositories listed
-
Deep Action Sequence Learning for Causal Shape Transformation17 May 2016 0 repositories listed
-
Real-Time Web Scale Event Summarization Using Sequential Decision Making12 May 2016 0 repositories listed
-
Stochastic Contextual Bandits with Known Reward Functions30 Apr 2016 0 repositories listed
-
Deep Learning for Reward Design to Improve Monte Carlo Tree Search in ATARI Games24 Apr 2016 0 repositories listed
-
Optimal Sensing via Multi-armed Bandit Relaxations in Mixed Observability Domains15 Mar 2016 0 repositories listed
-
PAC Reinforcement Learning with Rich Observations8 Feb 2016 0 repositories listed
-
Risk-Constrained Reinforcement Learning with Percentile Risk Criteria5 Dec 2015 0 repositories listed
-
Reuse of Neural Modules for General Video Game Playing4 Dec 2015 0 repositories listed
-
Bandits with Unobserved Confounders: A Causal Approach1 Dec 2015 0 repositories listed
-
Reinforcement Learning Applied to an Electric Water Heater: From Theory to Practice29 Nov 2015 0 repositories listed
-
Solving Transition-Independent Multi-agent MDPs with Sparse Interactions (Extended version)29 Nov 2015 0 repositories listed
-
The Knowledge Gradient with Logistic Belief Models for Binary Classification8 Oct 2015 0 repositories listed
-
Two Phase Q-learning for Bidding-based Vehicle Sharing29 Sep 2015 0 repositories listed
-
Optimization of anemia treatment in hemodialysis patients via reinforcement learning14 Sep 2015 0 repositories listed
-
Learning Efficient Representations for Reinforcement Learning28 Aug 2015 0 repositories listed
-
Experimental analysis of data-driven control for a building heating system13 Jul 2015 0 repositories listed
-
Utility-based Dueling Bandits as a Partial Monitoring Game10 Jul 2015 0 repositories listed
-
Hands-on Learning to Search for Structured Prediction1 May 2015 0 repositories listed
-
Global Bandits29 Mar 2015 0 repositories listed
-
Second-order Quantile Methods for Experts and Combinatorial Games27 Feb 2015 0 repositories listed
-
Fairness in Multi-Agent Sequential Decision-Making1 Dec 2014 0 repositories listed
-
Active Sensing as Bayes-Optimal Sequential Decision Making9 Aug 2014 0 repositories listed
-
Chasing Ghosts: Competing with Stateful Policies29 Jul 2014 0 repositories listed
-
Algorithms for CVaR Optimization in MDPs12 Jun 2014 0 repositories listed
-
Proximal Reinforcement Learning: A New Theory of Sequential Decision Making in Primal-Dual Spaces26 May 2014 0 repositories listed
-
Variance-Constrained Actor-Critic Algorithms for Discounted and Average Reward MDPs25 Mar 2014 0 repositories listed
-
A Survey of Multi-Objective Sequential Decision-Making4 Feb 2014 0 repositories listed
-
Exploiting Model Equivalences for Solving Interactive Dynamic Influence Diagrams18 Jan 2014 0 repositories listed
-
Non-Deterministic Policies in Markovian Decision Processes16 Jan 2014 0 repositories listed
-
Online Planning Algorithms for POMDPs15 Jan 2014 0 repositories listed
-
Actor-Critic Algorithms for Risk-Sensitive MDPs1 Dec 2013 0 repositories listed