Browse State-of-the-Art › Thompson Sampling › Papers, page 7
Thompson Sampling
Papers archive 2025-07-28
archive papers tagged: 655 · with a code link: 135 · where Syntology ran a sample: 30 (23 with a run with no instrument failure, 7 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (30 of 655 tagged: 23 with a run with no instrument failure, 7 where every run was a failure of Syntology's instrument)
Page 7 of 7: papers 601 to 655 of 655, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Adaptive Rate of Convergence of Thompson Sampling for Gaussian Process Optimization18 May 2017 0 repositories listed
-
Context Attentive Bandits: Contextual Bandit with Restricted Context10 May 2017 0 repositories listed
-
Multi-dueling Bandits with Dependent Arms29 Apr 2017 0 repositories listed
-
Time-Sensitive Bandit Learning and Satisficing Thompson Sampling28 Apr 2017 0 repositories listed
-
Efficient Benchmarking of NLP APIs using Multi-armed Bandits1 Apr 2017 0 repositories listed
-
Thompson Sampling for Linear-Quadratic Control Problems27 Mar 2017 0 repositories listed
-
Horde of Bandits using Gaussian Markov Random Fields7 Mar 2017 0 repositories listed
-
QoS-Aware Multi-Armed Bandits28 Feb 2017 0 repositories listed
-
Thompson Sampling For Stochastic Bandits with Graph Feedback16 Jan 2017 0 repositories listed
-
Estimating Quality in Multi-Objective Bandits Optimization4 Jan 2017 0 repositories listed
-
Exploration for Multi-task Reinforcement Learning with Deep Generative Models29 Nov 2016 0 repositories listed
-
Nonparametric General Reinforcement Learning28 Nov 2016 0 repositories listed
-
Linear Thompson Sampling Revisited20 Nov 2016 0 repositories listed
-
Unimodal Thompson Sampling for Graph-Structured Arms17 Nov 2016 0 repositories listed
-
The End of Optimism? An Asymptotic Analysis of Finite-Armed Linear Bandits14 Oct 2016 0 repositories listed
-
A Formal Solution to the Grain of Truth Problem16 Sep 2016 0 repositories listed
-
BBQ-Networks: Efficient Exploration in Deep Reinforcement Learning for Task-Oriented Dialogue Systems17 Aug 2016 0 repositories listed
-
Human collective intelligence as distributed Bayesian inference5 Aug 2016 0 repositories listed
-
Asymptotically Optimal Algorithms for Budgeted Multiple Play Bandits30 Jun 2016 0 repositories listed
-
Online Algorithms For Parameter Mean And Variance Estimation In Dynamic Regression Models18 May 2016 0 repositories listed
-
Linear Bandit algorithms using the Bootstrap4 May 2016 0 repositories listed
-
An Unbiased Data Collection and Content Exploitation/Exploration Strategy for Personalization12 Apr 2016 0 repositories listed
-
A sequential Monte Carlo approach to Thompson sampling for Bayesian optimization1 Apr 2016 0 repositories listed
-
Optimal Recommendation to Users that React: Online Learning for a Class of POMDPs30 Mar 2016 0 repositories listed
-
Thompson Sampling is Asymptotically Optimal in General Environments25 Feb 2016 0 repositories listed
-
Convolutional Monte Carlo Rollouts in Go10 Dec 2015 0 repositories listed
-
Efficient Thompson Sampling for Online Matrix-Factorization Recommendation1 Dec 2015 0 repositories listed
-
Regret Analysis of the Finite-Horizon Gittins Index Strategy for Multi-Armed Bandits18 Nov 2015 0 repositories listed
-
TSEB: More Efficient Thompson Sampling for Policy Learning10 Oct 2015 0 repositories listed
-
Bootstrapped Thompson Sampling and Deep Exploration1 Jul 2015 0 repositories listed
-
On the Prior Sensitivity of Thompson Sampling10 Jun 2015 0 repositories listed
-
Belief Flows of Robust Online Learning26 May 2015 0 repositories listed
-
Thompson Sampling for Budgeted Multi-armed Bandits1 May 2015 0 repositories listed
-
Evaluation of Explore-Exploit Policies in Multi-result Ranking Systems28 Apr 2015 0 repositories listed
-
A Note on Information-Directed Sampling and Thompson Sampling24 Mar 2015 0 repositories listed
-
Bandit Convex Optimization: sqrt{T} Regret in One Dimension23 Feb 2015 0 repositories listed
-
Thompson sampling with the online bootstrap15 Oct 2014 0 repositories listed
-
Freshness-Aware Thompson Sampling29 Sep 2014 0 repositories listed
-
Towards Optimal Algorithms for Prediction with Expert Advice10 Sep 2014 0 repositories listed
-
Thompson Sampling for Learning Parameterized Markov Decision Processes29 Jun 2014 0 repositories listed
-
Efficient Learning in Large-Scale Combinatorial Semi-Bandits28 Jun 2014 0 repositories listed
-
An Information-Theoretic Analysis of Thompson Sampling21 Mar 2014 0 repositories listed
-
Better Optimism By Bayes: Adaptive Planning with Rich Models9 Feb 2014 0 repositories listed
-
Bayesian Mixture Modelling and Inference based Thompson Sampling in Monte-Carlo Tree Search1 Dec 2013 0 repositories listed
-
Eluder Dimension and the Sample Complexity of Optimistic Exploration1 Dec 2013 0 repositories listed
-
Thompson Sampling for Complex Bandit Problems3 Nov 2013 0 repositories listed
-
Thompson Sampling for Online Learning with Linear Experts3 Nov 2013 0 repositories listed
-
Generalized Thompson Sampling for Contextual Bandits27 Oct 2013 0 repositories listed
-
Thompson Sampling in Dynamic Systems for Contextual Bandit Problems17 Oct 2013 0 repositories listed
-
Thompson Sampling for 1-Dimensional Exponential Family Bandits12 Jul 2013 0 repositories listed
-
Cover Tree Bayesian Reinforcement Learning8 May 2013 0 repositories listed
-
Prior-free and prior-dependent regret bounds for Thompson Sampling21 Apr 2013 0 repositories listed
-
Exploiting correlation and budget constraints in Bayesian multi-armed bandit optimization27 Mar 2013 0 repositories listed
-
Learning to Optimize Via Posterior Sampling11 Jan 2013 0 repositories listed
-
An Empirical Evaluation of Thompson Sampling1 Dec 2011 0 repositories listed