Browse State-of-the-Art › Multi-Armed Bandits › Papers, page 8
Multi-Armed Bandits
Papers archive 2025-07-28
archive papers tagged: 1,262 · with a code link: 253 · where Syntology ran a sample: 55 (44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (55 of 1,262 tagged: 44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 8 of 13: papers 701 to 800 of 1,262, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Breaking the √(T) Barrier: Instance-Independent Logarithmic Regret in Stochastic Contextual Linear Bandits19 May 2022 0 repositories listed
-
Slowly Changing Adversarial Bandit Algorithms are Efficient for Discounted MDPs18 May 2022 0 repositories listed
-
Semi-Parametric Contextual Bandits with Graph-Laplacian Regularization17 May 2022 0 repositories listed
-
From Dirichlet to Rubin: Optimistic Exploration in RL without Bonuses16 May 2022 0 repositories listed
-
Nearly Optimal Algorithms for Linear Contextual Bandits with Adversarial Corruptions13 May 2022 0 repositories listed
-
A Survey of Risk-Aware Multi-Armed Bandits12 May 2022 0 repositories listed
-
Federated Multi-Armed Bandits Under Byzantine Attacks9 May 2022 0 repositories listed
-
Selectively Contextual Bandits9 May 2022 0 repositories listed
-
Multi-Player Multi-Armed Bandits with Finite Shareable Resources Arms: Learning Algorithms & Applications28 Apr 2022 0 repositories listed
-
Rate-Constrained Remote Contextual Bandits26 Apr 2022 0 repositories listed
-
Worst-case Performance of Greedy Policies in Bandits with Imperfect Context Observations10 Apr 2022 0 repositories listed
-
Stochastic Multi-armed Bandits with Non-stationary Rewards Generated by a Linear Dynamical System6 Apr 2022 0 repositories listed
-
Strategies for Safe Multi-Armed Bandits with Logarithmic Regret and Risk1 Apr 2022 0 repositories listed
-
Flexible and Efficient Contextual Bandits with Heterogeneous Treatment Effect Oracles30 Mar 2022 0 repositories listed
-
Best Arm Identification in Restless Markov Multi-Armed Bandits29 Mar 2022 0 repositories listed
-
On Kernelized Multi-Armed Bandits with Constraints29 Mar 2022 0 repositories listed
-
Modeling Attrition in Recommender Systems with Departing Bandits25 Mar 2022 0 repositories listed
-
Approximate Function Evaluation via Multi-Armed Bandits18 Mar 2022 0 repositories listed
-
Reinforced Meta Active Learning9 Mar 2022 0 repositories listed
-
Reward-Biased Maximum Likelihood Estimation for Neural Contextual Bandits8 Mar 2022 0 repositories listed
-
PAC-Bayesian Lifelong Learning For Multi-Armed Bandits7 Mar 2022 0 repositories listed
-
Restless Multi-Armed Bandits under Exogenous Global Markov Process28 Feb 2022 0 repositories listed
-
Federated Online Sparse Decision Making27 Feb 2022 0 repositories listed
-
The Pareto Frontier of Instance-Dependent Guarantees in Multi-Player Multi-Armed Bandits with no Communication19 Feb 2022 0 repositories listed
-
Cost-Efficient Distributed Learning via Combinatorial Multi-Armed Bandits16 Feb 2022 0 repositories listed
-
Versatile Dueling Bandits: Best-of-both-World Analyses for Online Learning from Preferences14 Feb 2022 0 repositories listed
-
Shuffle Private Linear Contextual Bandits11 Feb 2022 0 repositories listed
-
Remote Contextual Bandits10 Feb 2022 0 repositories listed
-
Settling the Communication Complexity for Distributed Offline Reinforcement Learning10 Feb 2022 0 repositories listed
-
Smoothed Online Learning is as Easy as Statistical Learning9 Feb 2022 0 repositories listed
-
Budgeted Combinatorial Multi-Armed Bandits8 Feb 2022 0 repositories listed
-
Variance-Optimal Augmentation Logging for Counterfactual Evaluation in Contextual Bandits3 Feb 2022 0 repositories listed
-
Scalable Decision-Focused Learning in Restless Multi-Armed Bandits with Application to Maternal and Child Health2 Feb 2022 0 repositories listed
-
Efficient Algorithms for Learning to Control Bandits with Unobserved Contexts2 Feb 2022 0 repositories listed
-
Multi-armed Bandits for Link Configuration in Millimeter-wave Networks2 Feb 2022 0 repositories listed
-
Context Uncertainty in Contextual Bandits with Applications to Recommender Systems1 Feb 2022 0 repositories listed
-
Neural Collaborative Filtering Bandits via Meta Learning31 Jan 2022 0 repositories listed
-
Coordinated Attacks against Contextual Bandits: Fundamental Limits and Defense Mechanisms30 Jan 2022 0 repositories listed
-
Adaptive Best-of-Both-Worlds Algorithm for Heavy-Tailed Multi-Armed Bandits28 Jan 2022 0 repositories listed
-
Networked Restless Multi-Armed Bandits for Mobile Interventions28 Jan 2022 0 repositories listed
-
Top-K Ranking Deep Contextual Bandits for Information Selection Systems28 Jan 2022 0 repositories listed
-
Learning Neural Contextual Bandits Through Perturbed Rewards24 Jan 2022 0 repositories listed
-
Occupancy Information Ratio: Infinite-Horizon, Information-Directed, Parameterized Policy Search21 Jan 2022 0 repositories listed
-
Semantic Parsing for Planning Goals as Constrained Combinatorial Contextual Bandits16 Jan 2022 0 repositories listed
-
Contextual Bandits for Advertising Campaigns: A Diffusion-Model Independent Approach (Extended Version)13 Jan 2022 0 repositories listed
-
Modelling Cournot Games as Multi-agent Multi-armed Bandits1 Jan 2022 0 repositories listed
-
Safe Linear Leveling Bandits13 Dec 2021 0 repositories listed
-
Stochastic differential equations for limiting description of UCB rule for Gaussian multi-armed bandits13 Dec 2021 0 repositories listed
-
Privacy Amplification via Shuffling for Linear Contextual Bandits11 Dec 2021 0 repositories listed
-
Efficient Action Poisoning Attacks on Linear Contextual Bandits10 Dec 2021 0 repositories listed
-
Best Arm Identification under Additive Transfer Bandits8 Dec 2021 0 repositories listed
-
Contextual Bandit Applications in Customer Support Bot6 Dec 2021 0 repositories listed
-
On Submodular Contextual Bandits3 Dec 2021 0 repositories listed
-
Asymptotically Best Causal Effect Identification with Multi-Armed Bandits1 Dec 2021 0 repositories listed
-
Bandits with Knapsacks beyond the Worst Case1 Dec 2021 0 repositories listed
-
Multi-Armed Bandits with Bounded Arm-Memory: Near-Optimal Guarantees for Best-Arm Identification and Regret Minimization1 Dec 2021 0 repositories listed
-
Optimal Algorithms for Stochastic Contextual Preference Bandits1 Dec 2021 0 repositories listed
-
Online Fair Revenue Maximizing Cake Division with Non-Contiguous Pieces in Adversarial Bandits29 Nov 2021 0 repositories listed
-
Decentralized Upper Confidence Bound Algorithms for Homogeneous Multi-Agent Multi-Armed Bandits22 Nov 2021 0 repositories listed
-
Offline Contextual Bandits for Wireless Network Optimization11 Nov 2021 0 repositories listed
-
An Instance-Dependent Analysis for the Cooperative Multi-Player Multi-Armed Bandit8 Nov 2021 0 repositories listed
-
Universal and data-adaptive algorithms for model selection in linear contextual bandits8 Nov 2021 0 repositories listed
-
Privacy-Preserving Communication-Efficient Federated Multi-Armed Bandits2 Nov 2021 0 repositories listed
-
Bandits Don’t Follow Rules: Balancing Multi-Facet Machine Translation with Multi-Armed Bandits1 Nov 2021 0 repositories listed
-
Decentralized Cooperative Reinforcement Learning with Hierarchical Information Structure1 Nov 2021 0 repositories listed
-
Federated Linear Contextual Bandits27 Oct 2021 0 repositories listed
-
Linear Contextual Bandits with Adversarial Corruptions25 Oct 2021 0 repositories listed
-
The Pareto Frontier of model selection for general Contextual Bandits25 Oct 2021 0 repositories listed
-
Analysis of Thompson Sampling for Partially Observable Contextual Multi-Armed Bandits23 Oct 2021 0 repositories listed
-
Dynamic pricing and assortment under a contextual MNL demand19 Oct 2021 0 repositories listed
-
Stateful Offline Contextual Policy Evaluation and Learning19 Oct 2021 0 repositories listed
-
Achieving the Pareto Frontier of Regret Minimization and Best Arm Identification in Multi-Armed Bandits16 Oct 2021 0 repositories listed
-
Almost Optimal Batch-Regret Tradeoff for Batch Linear Contextual Bandits15 Oct 2021 0 repositories listed
-
Bandits Don't Follow Rules: Balancing Multi-Facet Machine Translation with Multi-Armed Bandits13 Oct 2021 0 repositories listed
-
Query-Reward Tradeoffs in Multi-Armed Bandits12 Oct 2021 0 repositories listed
-
Deep Upper Confidence Bound Algorithm for Contextual Bandit Ranking of Information Selection8 Oct 2021 0 repositories listed
-
A Model Selection Approach for Corruption Robust Reinforcement Learning7 Oct 2021 0 repositories listed
-
Feel-Good Thompson Sampling for Contextual Bandits and Reinforcement Learning2 Oct 2021 0 repositories listed
-
Asymptotic Performance of Thompson Sampling in the Batched Multi-Armed Bandits1 Oct 2021 0 repositories listed
-
Batched Thompson Sampling1 Oct 2021 0 repositories listed
-
Adapting Bandit Algorithms for Settings with Sequentially Available Arms30 Sep 2021 0 repositories listed
-
Batched Bandits with Crowd Externalities29 Sep 2021 0 repositories listed
-
Causal Contextual Bandits with Targeted Interventions29 Sep 2021 0 repositories listed
-
Expected Improvement-based Contextual Bandits29 Sep 2021 0 repositories listed
-
Regularized-OFU: an efficient algorithm for general contextual bandit with optimization oracles29 Sep 2021 0 repositories listed
-
Risk averse non-stationary multi-armed bandits28 Sep 2021 0 repositories listed
-
Robust Generalization of Quadratic Neural Networks via Function Identification22 Sep 2021 0 repositories listed
-
Generalized Translation and Scale Invariant Online Algorithm for Adversarial Multi-Armed Bandits19 Sep 2021 0 repositories listed
-
Field Study in Deploying Restless Multi-Armed Bandits: Assisting Non-Profits in Improving Maternal and Child Health16 Sep 2021 0 repositories listed
-
Exploiting Heterogeneity in Robust Federated Best-Arm Identification13 Sep 2021 0 repositories listed
-
Improved Algorithms for Misspecified Linear Markov Decision Processes12 Sep 2021 0 repositories listed
-
Best-Arm Identification in Correlated Multi-Armed Bandits10 Sep 2021 0 repositories listed
-
Online Learning for Cooperative Multi-Player Multi-Armed Bandits7 Sep 2021 0 repositories listed
-
Max-Utility Based Arm Selection Strategy For Sequential Query Recommendations31 Aug 2021 0 repositories listed
-
No DBA? No regret! Multi-armed bandits for index tuning of analytical and HTAP workloads with provable guarantees23 Aug 2021 0 repositories listed
-
Batched Thompson Sampling for Multi-Armed Bandits15 Aug 2021 0 repositories listed
-
Metadata-based Multi-Task Bandits with Bayesian Hierarchical Models13 Aug 2021 0 repositories listed
-
Regret Analysis of Learning-Based MPC with Partially-Unknown Cost Function4 Aug 2021 0 repositories listed
-
Indexability and Rollout Policy for Multi-State Partially Observable Restless Bandits30 Jul 2021 0 repositories listed
-
Combining Online Learning and Offline Learning for Contextual Bandits with Deficient Support24 Jul 2021 0 repositories listed