Browse State-of-the-Art › Multi-Armed Bandits › Papers, page 11
Multi-Armed Bandits
Papers archive 2025-07-28
archive papers tagged: 1,262 · with a code link: 253 · where Syntology ran a sample: 55 (44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (55 of 1,262 tagged: 44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 11 of 13: papers 1,001 to 1,100 of 1,262, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Distributionally Robust Policy Evaluation and Learning in Offline Contextual Bandits1 Jan 2020 0 repositories listed
-
Exploration Through Bias: Revisiting Biased Maximum Likelihood Estimation in Stochastic Multi-Armed Bandits1 Jan 2020 0 repositories listed
-
Gradient-free Online Learning in Continuous Games with Delayed Rewards1 Jan 2020 0 repositories listed
-
Fair Contextual Multi-Armed Bandits: Theory and Experiments13 Dec 2019 0 repositories listed
-
Sublinear Optimal Policy Value Estimation in Contextual Bandits12 Dec 2019 0 repositories listed
-
Epsilon-Best-Arm Identification in Pay-Per-Reward Multi-Armed Bandits1 Dec 2019 0 repositories listed
-
Learning in Generalized Linear Contextual Bandits with Stochastic Delays1 Dec 2019 0 repositories listed
-
Nonparametric Contextual Bandits in Metric Spaces with Unknown Metric1 Dec 2019 0 repositories listed
-
Surrogate Objectives for Batch Policy Optimization in One-step Decision Making1 Dec 2019 0 repositories listed
-
Contextual Combinatorial Conservative Bandits26 Nov 2019 0 repositories listed
-
Automatic Ensemble Learning for Online Influence Maximization25 Nov 2019 0 repositories listed
-
Corruption-robust exploration in episodic reinforcement learning20 Nov 2019 0 repositories listed
-
Contextual Bandits Evolving Over Finite Time14 Nov 2019 0 repositories listed
-
Unreliable Multi-Armed Bandits: A Novel Approach to Recommendation Systems14 Nov 2019 0 repositories listed
-
Triply Robust Off-Policy Evaluation13 Nov 2019 0 repositories listed
-
Incentivized Exploration for Multi-Armed Bandits under Reward Drift12 Nov 2019 0 repositories listed
-
Problem Dependent Reinforcement Learning Bounds Which Can Identify Bandit Structure in MDPs3 Nov 2019 0 repositories listed
-
Trend-responsive User Segmentation Enabling Traceable Publishing Insights. A Case Study of a Real-world Large-scale News Recommendation System28 Oct 2019 0 repositories listed
-
BanditRank: Learning to Rank Using Contextual Bandits23 Oct 2019 0 repositories listed
-
Decentralized Heterogeneous Multi-Player Multi-Armed Bandits with Non-Zero Rewards on Collisions21 Oct 2019 0 repositories listed
-
Multi-User MABs with User Dependent Rewards for Uncoordinated Spectrum Access21 Oct 2019 0 repositories listed
-
Adaptive Exploration in Linear Contextual Bandit15 Oct 2019 0 repositories listed
-
An Optimal Algorithm for Adversarial Bandits with Arbitrary Delays14 Oct 2019 0 repositories listed
-
Regret Bounds for Batched Bandits11 Oct 2019 0 repositories listed
-
Privacy-Preserving Multi-Party Contextual Bandits11 Oct 2019 0 repositories listed
-
Social Learning in Multi Agent Multi Armed Bandits4 Oct 2019 0 repositories listed
-
Decision Automation for Electric Power Network Recovery1 Oct 2019 0 repositories listed
-
An Optimal Algorithm for Multiplayer Multi-Armed Bandits28 Sep 2019 0 repositories listed
-
Learning Effective Exploration Strategies For Contextual Bandits25 Sep 2019 0 repositories listed
-
NeuralUCB: Contextual Bandits with Neural Network-Based Exploration25 Sep 2019 0 repositories listed
-
AutoML for Contextual Bandits7 Sep 2019 0 repositories listed
-
A Near-Optimal Change-Detection Based Algorithm for Piecewise-Stationary Combinatorial Semi-Bandits27 Aug 2019 0 repositories listed
-
Nonparametric Contextual Bandits in an Unknown Metric Space3 Aug 2019 0 repositories listed
-
Parameterized Exploration13 Jul 2019 0 repositories listed
-
Productization Challenges of Contextual Multi-Armed Bandits10 Jul 2019 0 repositories listed
-
Individual Regret in Cooperative Nonstochastic Multi-Armed Bandits7 Jul 2019 0 repositories listed
-
Exploration Through Reward Biasing: Reward-Biased Maximum Likelihood Estimation for Stochastic Multi-Armed Bandits2 Jul 2019 0 repositories listed
-
Multi-Armed Bandits with Fairness Constraints for Distributing Resources to Human Teammates30 Jun 2019 0 repositories listed
-
Learning in Restless Multi-Armed Bandits via Adaptive Arm Sequencing Rules19 Jun 2019 0 repositories listed
-
Online Allocation and Pricing: Constant Regret via Bellman Inequalities14 Jun 2019 0 repositories listed
-
Bootstrapping Upper Confidence Bound12 Jun 2019 0 repositories listed
-
Competing Bandits in Matching Markets12 Jun 2019 0 repositories listed
-
Beam Learning -- Using Machine Learning for Finding Beam Directions11 Jun 2019 0 repositories listed
-
Stochastic Neural Network with Kronecker Flow10 Jun 2019 0 repositories listed
-
Distribution-dependent and Time-uniform Bounds for Piecewise i.i.d Bandits30 May 2019 0 repositories listed
-
Equipping Experts/Bandits with Long-term Memory30 May 2019 0 repositories listed
-
Multi-Objective Generalized Linear Bandits30 May 2019 0 repositories listed
-
Rarely-switching linear bandits: optimization of causal effects for the real world30 May 2019 0 repositories listed
-
Differential Privacy for Multi-armed Bandits: What Is It and What Is Its Cost?29 May 2019 0 repositories listed
-
Top-k Combinatorial Bandits with Full-Bandit Feedback28 May 2019 0 repositories listed
-
Achieving Fairness in Stochastic Multi-armed Bandit Problem27 May 2019 0 repositories listed
-
Are sample means in multi-armed bandits positively or negatively biased?27 May 2019 0 repositories listed
-
OSOM: A simultaneously optimal algorithm for multi-armed and linear contextual bandits24 May 2019 0 repositories listed
-
Data Poisoning Attacks on Stochastic Bandits16 May 2019 0 repositories listed
-
Lessons from Contextual Bandit Learning in a Customer Support Bot6 May 2019 0 repositories listed
-
Tight Regret Bounds for Infinite-armed Linear Contextual Bandits4 May 2019 0 repositories listed
-
Meta-learners' learning dynamics are unlike learners'3 May 2019 0 repositories listed
-
Non-Stochastic Multi-Player Multi-Armed Bandits: Optimal Rate With Collision Information, Sublinear Without28 Apr 2019 0 repositories listed
-
Constrained Restless Bandits for Dynamic Scheduling in Cyber-Physical Systems18 Apr 2019 0 repositories listed
-
Distributed Bandit Learning: Near-Optimal Regret with Efficient Communication12 Apr 2019 0 repositories listed
-
Collaborative Learning with Limited Interaction: Tight Bounds for Distributed Exploration in Multi-Armed Bandits5 Apr 2019 0 repositories listed
-
A Survey on Practical Applications of Multi-Armed and Contextual Bandits2 Apr 2019 0 repositories listed
-
Nearly Minimax-Optimal Regret for Linearly Parameterized Bandits30 Mar 2019 0 repositories listed
-
Meta-Learning surrogate models for sequential decision making28 Mar 2019 0 repositories listed
-
Contextual Bandits with Random Projection20 Mar 2019 0 repositories listed
-
Perturbed-History Exploration in Stochastic Multi-Armed Bandits26 Feb 2019 0 repositories listed
-
Better Algorithms for Stochastic Bandits with Adversarial Corruptions22 Feb 2019 0 repositories listed
-
AdaLinUCB: Opportunistic Learning for Contextual Bandits20 Feb 2019 0 repositories listed
-
Contextual Bandits with Continuous Actions: Smoothing, Zooming, and Adapting5 Feb 2019 0 repositories listed
-
A New Algorithm for Non-stationary Contextual Bandits: Efficient, Optimal, and Parameter-free3 Feb 2019 0 repositories listed
-
Randomized Allocation with Nonparametric Estimation for Contextual Multi-Armed Bandits with Delayed Rewards3 Feb 2019 0 repositories listed
-
On the bias, risk and consistency of sample means in multi-armed bandits2 Feb 2019 0 repositories listed
-
Target Tracking for Contextual Bandits: Application to Demand Side Management28 Jan 2019 0 repositories listed
-
Almost Boltzmann Exploration25 Jan 2019 0 repositories listed
-
Deep Neural Linear Bandits: Overcoming Catastrophic Forgetting through Likelihood Matching24 Jan 2019 0 repositories listed
-
PAC Identification of Many Good Arms in Stochastic Multi-Armed Bandits24 Jan 2019 0 repositories listed
-
Regret Minimisation in Multi-Armed Bandits Using Bounded Arm Memory24 Jan 2019 0 repositories listed
-
Parallel Contextual Bandits in Wireless Handover Optimization21 Jan 2019 0 repositories listed
-
Imitation-Regularized Offline Learning15 Jan 2019 0 repositories listed
-
Concentration bounds for CVaR estimation: The cases of light-tailed and heavy-tailed distributions4 Jan 2019 0 repositories listed
-
Multi-player Multi-armed Bandits for Stable Allocation in Heterogeneous Ad-Hoc Networks24 Dec 2018 0 repositories listed
-
Human-AI Learning Performance in Multi-Armed Bandits21 Dec 2018 0 repositories listed
-
Generalizable Meta-Heuristic based on Temporal Estimation of Rewards for Large Scale Blackbox Optimization17 Dec 2018 0 repositories listed
-
Balanced Linear Contextual Bandits15 Dec 2018 0 repositories listed
-
ADARES: Adaptive Resource Management for Virtual Machines5 Dec 2018 0 repositories listed
-
A Bandit Approach to Sequential Experimental Design with False Discovery Control1 Dec 2018 0 repositories listed
-
Contextual Combinatorial Multi-armed Bandits with Volatile Arms and Submodular Reward1 Dec 2018 0 repositories listed
-
Why so gloomy? A Bayesian explanation of human pessimism bias in the multi-armed bandit task1 Dec 2018 0 repositories listed
-
Stochastic Top-K Subset Bandits with Linear Space and Non-Linear Feedback29 Nov 2018 0 repositories listed
-
Adversarial Bandits with Knapsacks28 Nov 2018 0 repositories listed
-
Kernel-based Multi-Task Contextual Bandits in Cellular Network Configuration27 Nov 2018 0 repositories listed
-
Rotting bandits are not harder than stochastic ones27 Nov 2018 0 repositories listed
-
Bandits with Temporal Stochastic Constraints22 Nov 2018 0 repositories listed
-
Best Arm Identification in Linked Bandits19 Nov 2018 0 repositories listed
-
Decentralized Exploration in Multi-Armed Bandits -- Extended version19 Nov 2018 0 repositories listed
-
Sample complexity of partition identification using multi-armed bandits14 Nov 2018 0 repositories listed
-
Garbage In, Reward Out: Bootstrapping Exploration in Multi-Armed Bandits13 Nov 2018 0 repositories listed
-
Multi-armed Bandits with Compensation5 Nov 2018 0 repositories listed
-
Online learning with feedback graphs and switching costs23 Oct 2018 0 repositories listed
-
Simple Regret Minimization for Contextual Bandits17 Oct 2018 0 repositories listed