Browse State-of-the-Art › Multi-Armed Bandits › Papers, page 12
Multi-Armed Bandits
Papers archive 2025-07-28
archive papers tagged: 1,262 · with a code link: 253 · where Syntology ran a sample: 55 (44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (55 of 1,262 tagged: 44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 12 of 13: papers 1,101 to 1,200 of 1,262, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Fighting Contextual Bandits with Stochastic Smoothing11 Oct 2018 0 repositories listed
-
Regularized Contextual Bandits11 Oct 2018 0 repositories listed
-
Contextual Multi-Armed Bandits for Causal Marketing2 Oct 2018 0 repositories listed
-
Thompson Sampling Algorithms for Cascading Bandits2 Oct 2018 0 repositories listed
-
Contextual Bandits with Cross-learning25 Sep 2018 0 repositories listed
-
Multi-Player Bandits: A Trekking Approach17 Sep 2018 0 repositories listed
-
Diversity-Driven Selection of Exploration Strategies in Multi-Armed Bandits23 Aug 2018 0 repositories listed
-
Data Poisoning Attacks in Contextual Bandits17 Aug 2018 0 repositories listed
-
Preference-based Online Learning with Dueling Bandits: A Survey30 Jul 2018 0 repositories listed
-
Deep Contextual Multi-armed Bandits25 Jul 2018 0 repositories listed
-
Tsallis-INF: An Optimal Algorithm for Stochastic and Adversarial Bandits19 Jul 2018 0 repositories listed
-
Linear Bandits with Stochastic Delayed Feedback5 Jul 2018 0 repositories listed
-
Multi-User Multi-Armed Bandits for Uncoordinated Spectrum Access2 Jul 2018 0 repositories listed
-
Learning to Coordinate with Coordination Graphs in Repeated Single-Stage Multi-Agent Decision Problems1 Jul 2018 0 repositories listed
-
Contextual bandits with surrogate losses: Margin bounds and efficient algorithms28 Jun 2018 0 repositories listed
-
Greybox fuzzing as a contextual bandits problem11 Jun 2018 0 repositories listed
-
Finding the bandit in a graph: Sequential search-and-stop6 Jun 2018 0 repositories listed
-
Mitigating Bias in Adaptive Data Gathering via Differential Privacy6 Jun 2018 0 repositories listed
-
A General Framework for Bandit Problems Beyond Cumulative Objectives4 Jun 2018 0 repositories listed
-
The Externalities of Exploration and How Data Diversity Helps Exploitation1 Jun 2018 0 repositories listed
-
Multi-Statistic Approximate Bayesian Computation with Multi-Armed Bandits22 May 2018 0 repositories listed
-
PG-TS: Improved Thompson Sampling for Logistic Contextual Bandits18 May 2018 0 repositories listed
-
Delegating via Quitting Games20 Apr 2018 0 repositories listed
-
Combining Difficulty Ranking with Multi-Armed Bandits to Sequence Educational Content14 Apr 2018 0 repositories listed
-
Best arm identification in multi-armed bandits with delayed feedback29 Mar 2018 0 repositories listed
-
What Doubling Tricks Can and Can't Do for Multi-Armed Bandits19 Mar 2018 0 repositories listed
-
Multi-Armed Bandits for Correlated Markovian Environments with Smoothed Reward Feedback11 Mar 2018 0 repositories listed
-
Online learning over a finite action set with limited switching5 Mar 2018 0 repositories listed
-
Practical Contextual Bandits with Regression Oracles3 Mar 2018 0 repositories listed
-
The K-Nearest Neighbour UCB algorithm for multi-armed bandits with covariates1 Mar 2018 0 repositories listed
-
Regional Multi-Armed Bandits22 Feb 2018 0 repositories listed
-
Online Learning with an Unknown Fairness Metric20 Feb 2018 0 repositories listed
-
Multi-Armed Bandits on Partially Revealed Unit Interval Graphs12 Feb 2018 0 repositories listed
-
Policy Gradients for Contextual Recommendations12 Feb 2018 0 repositories listed
-
Make the Minority Great Again: First-Order Regret Bound for Contextual Bandits9 Feb 2018 0 repositories listed
-
Nonparametric Stochastic Contextual Bandits5 Jan 2018 0 repositories listed
-
Contextual memory bandit for pro-active dialog engagement1 Jan 2018 0 repositories listed
-
Active Search for High Recall: a Non-Stationary Extension of Thompson Sampling27 Dec 2017 0 repositories listed
-
Stochastic Multi-armed Bandits in Constant Space25 Dec 2017 0 repositories listed
-
Gaussian Process bandits with adaptive discretization5 Dec 2017 0 repositories listed
-
A KL-LUCB algorithm for Large-Scale Crowdsourcing1 Dec 2017 0 repositories listed
-
Online Learning via the Differential Privacy Lens27 Nov 2017 0 repositories listed
-
Customized Nonlinear Bandits for Online Response Selection in Neural Conversation Models22 Nov 2017 0 repositories listed
-
Estimation Considerations in Contextual Bandits19 Nov 2017 0 repositories listed
-
Budget-Constrained Multi-Armed Bandits with Multiple Plays16 Nov 2017 0 repositories listed
-
Skyline Identification in Multi-Armed Bandits12 Nov 2017 0 repositories listed
-
Small-loss bounds for online learning with partial information9 Nov 2017 0 repositories listed
-
Multi-Player Bandits Revisited7 Nov 2017 0 repositories listed
-
Sparsity, variance and curvature in multi-armed bandits3 Nov 2017 0 repositories listed
-
Multi-Armed Bandits with Metric Movement Costs24 Oct 2017 0 repositories listed
-
Combinatorial Multi-armed Bandits for Real-Time Strategy Games13 Oct 2017 0 repositories listed
-
An Analysis of the Value of Information when Exploring Stochastic, Discrete Multi-Armed Bandits8 Oct 2017 0 repositories listed
-
Trend Detection based Regret Minimization for Bandit Problems15 Sep 2017 0 repositories listed
-
Optimal Learning for Sequential Decision Making for Expensive Cost Functions with Stochastic Binary Feedbacks13 Sep 2017 0 repositories listed
-
Ease.ml: Towards Multi-tenant Resource Sharing for Machine Learning Workloads24 Aug 2017 0 repositories listed
-
Efficient Contextual Bandits in Non-stationary Worlds5 Aug 2017 0 repositories listed
-
Reinforcement learning techniques for Outer Loop Link Adaptation in 4G/5G systems3 Aug 2017 0 repositories listed
-
Safety-Aware Algorithms for Adversarial Contextual Bandit1 Aug 2017 0 repositories listed
-
A Survey of Learning in Multiagent Environments: Dealing with Non-Stationarity28 Jul 2017 0 repositories listed
-
Nonlinear Sequential Accepts and Rejects for Identification of Top Arms in Stochastic Bandits9 Jul 2017 0 repositories listed
-
Efficient Reinforcement Learning via Initial Pure Exploration7 Jun 2017 0 repositories listed
-
Nearly Optimal Sampling Algorithms for Combinatorial Pure Exploration4 Jun 2017 0 repositories listed
-
Boltzmann Exploration Done Right29 May 2017 0 repositories listed
-
Combinatorial Multi-Armed Bandits with Filtered Feedback26 May 2017 0 repositories listed
-
Boundary Crossing Probabilities for General Exponential Families24 May 2017 0 repositories listed
-
Multi-Task Learning for Contextual Bandits24 May 2017 0 repositories listed
-
Combinatorial Semi-Bandits with Knapsacks23 May 2017 0 repositories listed
-
Practical Algorithms for Best-K Identification in Multi-Armed Bandits19 May 2017 0 repositories listed
-
Bandit Regret Scaling with the Effective Loss Range15 May 2017 0 repositories listed
-
Value Directed Exploration in Multi-Armed Bandits with Structured Priors12 Apr 2017 0 repositories listed
-
On Kernelized Multi-armed Bandits3 Apr 2017 0 repositories listed
-
Efficient Benchmarking of NLP APIs using Multi-armed Bandits1 Apr 2017 0 repositories listed
-
Selective Harvesting over Networks15 Mar 2017 0 repositories listed
-
Horde of Bandits using Gaussian Markov Random Fields7 Mar 2017 0 repositories listed
-
Provably Optimal Algorithms for Generalized Linear Contextual Bandits28 Feb 2017 0 repositories listed
-
QoS-Aware Multi-Armed Bandits28 Feb 2017 0 repositories listed
-
Rotting Bandits23 Feb 2017 0 repositories listed
-
Beyond the Hazard Rate: More Perturbation Algorithms for Adversarial Multi-armed Bandits17 Feb 2017 0 repositories listed
-
Learning to Use Learners' Advice16 Feb 2017 0 repositories listed
-
The Price of Differential Privacy For Online Learning27 Jan 2017 0 repositories listed
-
Active Search for Sparse Signals with Region Sensing2 Dec 2016 0 repositories listed
-
Multi-armed Bandits: Competing with Optimal Sequences1 Dec 2016 0 repositories listed
-
Bandit algorithms to emulate human decision making using probabilistic distortions30 Nov 2016 0 repositories listed
-
Fair Algorithms for Infinite and Contextual Bandits29 Oct 2016 0 repositories listed
-
Risk-Aware Algorithms for Adversarial Contextual Bandits17 Oct 2016 0 repositories listed
-
Exploration Potential16 Sep 2016 0 repositories listed
-
On Sequential Elimination Algorithms for Best-Arm Identification in Multi-Armed Bandits8 Sep 2016 0 repositories listed
-
On the Identification and Mitigation of Weaknesses in the Knowledge Gradient Policy for Multi-Armed Bandits20 Jul 2016 0 repositories listed
-
An optimal learning method for developing personalized treatment regimes6 Jul 2016 0 repositories listed
-
Making Contextual Decisions with Low Technical Debt13 Jun 2016 0 repositories listed
-
Contextual Bandits with Latent Confounders: An NMF Approach1 Jun 2016 0 repositories listed
-
Improved Regret Bounds for Oracle-Based Adversarial Contextual Bandits1 Jun 2016 0 repositories listed
-
Open Problem: Best Arm Identification: Almost Instance-Wise Optimality and the Gap Entropy Conjecture27 May 2016 0 repositories listed
-
Fairness in Learning: Classic and Contextual Bandits23 May 2016 0 repositories listed
-
Graph Clustering Bandits for Recommendation2 May 2016 0 repositories listed
-
Stochastic Contextual Bandits with Known Reward Functions30 Apr 2016 0 repositories listed
-
Latent Contextual Bandits and their Application to Personalized Recommendations for New Users22 Apr 2016 0 repositories listed
-
PAC Reinforcement Learning with Rich Observations8 Feb 2016 0 repositories listed
-
BISTRO: An Efficient Relaxation-Based Method for Contextual Bandits6 Feb 2016 0 repositories listed
-
Bandits meet Computer Architecture: Designing a Smartly-allocated Cache31 Jan 2016 0 repositories listed