Browse State-of-the-Art › Multi-Armed Bandits › Papers, page 13
Multi-Armed Bandits
Papers archive 2025-07-28
archive papers tagged: 1,262 · with a code link: 253 · where Syntology ran a sample: 55 (44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (55 of 1,262 tagged: 44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 13 of 13: papers 1,201 to 1,262 of 1,262, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Personalized Course Sequence Recommendations30 Dec 2015 0 repositories listed
-
On Top-k Selection in Multi-Armed Bandits and Hidden Bipartite Graphs1 Dec 2015 0 repositories listed
-
Algorithms for Differentially Private Multi-Armed Bandits27 Nov 2015 0 repositories listed
-
Regret Analysis of the Finite-Horizon Gittins Index Strategy for Multi-Armed Bandits18 Nov 2015 0 repositories listed
-
Context-Aware Bandits12 Oct 2015 0 repositories listed
-
Multi-armed Bandits with Application to 5G Small Cells2 Oct 2015 0 repositories listed
-
Sequential Design for Ranking Response Surfaces3 Sep 2015 0 repositories listed
-
Episodic Multi-armed Bandits4 Aug 2015 0 repositories listed
-
Linear Contextual Bandits with Knapsacks24 Jul 2015 0 repositories listed
-
Selecting the best system and multi-armed bandits16 Jul 2015 0 repositories listed
-
Upper-Confidence-Bound Algorithms for Active Learning in Multi-Armed Bandits16 Jul 2015 0 repositories listed
-
Scalable Discrete Sampling as a Multi-Armed Bandit Problem30 Jun 2015 0 repositories listed
-
An efficient algorithm for contextual bandits with knapsacks, and an extension to concave objectives10 Jun 2015 0 repositories listed
-
On Regret-Optimal Learning in Decentralized Multi-player Multi-armed Bandits4 May 2015 0 repositories listed
-
Thompson Sampling for Budgeted Multi-armed Bandits1 May 2015 0 repositories listed
-
Algorithms with Logarithmic or Sublinear Regret for Constrained Contextual Bandits27 Apr 2015 0 repositories listed
-
Regret vs. Communication: Distributed Stochastic Multi-Armed Bandits and Beyond14 Apr 2015 0 repositories listed
-
Global Bandits29 Mar 2015 0 repositories listed
-
Networked Stochastic Multi-Armed Bandits with Combinatorial Strategies20 Mar 2015 0 repositories listed
-
Learning to Search Better Than Your Teacher8 Feb 2015 0 repositories listed
-
Combinatorial Pure Exploration of Multi-Armed Bandits1 Dec 2014 0 repositories listed
-
Learning Multiple Tasks in Parallel with a Shared Annotator1 Dec 2014 0 repositories listed
-
Nonstochastic Multi-Armed Bandits with Graph-Structured Feedback30 Sep 2014 0 repositories listed
-
On Minimax Optimal Offline Policy Evaluation12 Sep 2014 0 repositories listed
-
Bandits Warm-up Cold Recommender Systems10 Jul 2014 0 repositories listed
-
Unimodal Bandits: Regret Lower Bounds and Optimal Algorithms20 May 2014 0 repositories listed
-
Lipschitz Bandits: Regret Lower Bounds and Optimal Algorithms19 May 2014 0 repositories listed
-
Reducing Dueling Bandits to Cardinal Bandits14 May 2014 0 repositories listed
-
Adaptive Contract Design for Crowdsourcing Markets: Bandit Algorithms for Repeated Principal-Agent Problems12 May 2014 0 repositories listed
-
Generalized Risk-Aversion in Stochastic Multi-Armed Bandits5 May 2014 0 repositories listed
-
Resourceful Contextual Bandits27 Feb 2014 0 repositories listed
-
Algorithms for multi-armed bandit problems25 Feb 2014 0 repositories listed
-
Exploration vs Exploitation vs Safety: Risk-averse Multi-Armed Bandits6 Jan 2014 0 repositories listed
-
lil' UCB : An Optimal Exploration Algorithm for Multi-Armed Bandits27 Dec 2013 0 repositories listed
-
Fundamental Limits of Online and Distributed Algorithms for Statistical Learning and Estimation14 Nov 2013 0 repositories listed
-
Distributed Exploration in Multi-Armed Bandits4 Nov 2013 0 repositories listed
-
Generalized Thompson Sampling for Contextual Bandits27 Oct 2013 0 repositories listed
-
Multi-Armed Bandits for Intelligent Tutoring Systems11 Oct 2013 0 repositories listed
-
Sequential Monte Carlo Bandits4 Oct 2013 0 repositories listed
-
Building Bridges: Viewing Active Learning from the Multi-Armed Bandit Lens26 Sep 2013 0 repositories listed
-
Finite-Time Analysis of Kernelised Contextual Bandits26 Sep 2013 0 repositories listed
-
Distributed Online Learning via Cooperative Contextual Bandits21 Aug 2013 0 repositories listed
-
Modeling Human Decision-making in Generalized Gaussian Multi-armed Bandits23 Jul 2013 0 repositories listed
-
Towards Distribution-Free Multi-Armed Bandits with Combinatorial Strategies20 Jul 2013 0 repositories listed
-
From Bandits to Experts: A Tale of Domination and Independence17 Jul 2013 0 repositories listed
-
On Finding the Largest Mean Among Many17 Jun 2013 0 repositories listed
-
Concentration bounds for temporal difference learning with linear function approximation: The case of batch data and uniform sampling11 Jun 2013 0 repositories listed
-
A Gang of Bandits4 Jun 2013 0 repositories listed
-
Dynamic Ad Allocation: Bandits with Budgets1 Jun 2013 0 repositories listed
-
Exponentiated Gradient LINUCB for Contextual Multi-Armed Bandits10 May 2013 0 repositories listed
-
Hierarchical Optimistic Region Selection driven by Curiosity1 Dec 2012 0 repositories listed
-
Risk-Aversion in Multi-armed Bandits1 Dec 2012 0 repositories listed
-
An Empirical Evaluation of Thompson Sampling1 Dec 2011 0 repositories listed
-
From Bandits to Experts: On the Value of Side-Observations1 Dec 2011 0 repositories listed
-
Multi-armed bandits on implicit metric spaces1 Dec 2011 0 repositories listed
-
PAC-Bayesian Analysis of Contextual Bandits1 Dec 2011 0 repositories listed
-
Dynamic Pricing with Limited Supply20 Aug 2011 0 repositories listed
-
Combinatorial Network Optimization with Unknown Variables: Multi-Armed Bandits with Linear Rewards22 Nov 2010 0 repositories listed
-
Contextual Bandits with Similarity Information23 Jul 2009 0 repositories listed
-
Mortal Multi-Armed Bandits1 Dec 2008 0 repositories listed
-
Learning diverse rankings with multi-armed bandits5 Jul 2008 0 repositories listed
-
The Epoch-Greedy Algorithm for Multi-armed Bandits with Side Information1 Dec 2007 0 repositories listed