Browse State-of-the-Art › Multi-Armed Bandits › Papers, page 7
Multi-Armed Bandits
Papers archive 2025-07-28
archive papers tagged: 1,262 · with a code link: 253 · where Syntology ran a sample: 55 (44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (55 of 1,262 tagged: 44 with a run with no instrument failure, 11 where every run was a failure of Syntology's instrument)
Page 7 of 13: papers 601 to 700 of 1,262, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Leveraging User-Triggered Supervision in Contextual Bandits7 Feb 2023 0 repositories listed
-
On Private and Robust Bandits6 Feb 2023 0 repositories listed
-
Multiplier Bootstrap-based Exploration3 Feb 2023 0 repositories listed
-
Randomized Greedy Learning for Non-monotone Stochastic Submodular Maximization Under Full-bandit Feedback2 Feb 2023 0 repositories listed
-
Stochastic Contextual Bandits with Long Horizon Rewards2 Feb 2023 0 repositories listed
-
Improved Algorithms for Multi-period Multi-class Packing Problems with Bandit Feedback31 Jan 2023 0 repositories listed
-
Quantum contextual bandits and recommender systems for quantum data31 Jan 2023 0 repositories listed
-
A Framework for Adapting Offline Algorithms to Solve Combinatorial Multi-Armed Bandit Problems with Bandit Feedback30 Jan 2023 0 repositories listed
-
Adversarial Attacks on Adversarial Bandits30 Jan 2023 0 repositories listed
-
Contextual Causal Bayesian Optimisation29 Jan 2023 0 repositories listed
-
Communication-Efficient Collaborative Regret Minimization in Multi-Armed Bandits26 Jan 2023 0 repositories listed
-
Banker Online Mirror Descent: A Universal Approach for Delayed Online Bandit Learning25 Jan 2023 0 repositories listed
-
Quantum Heavy-tailed Bandits23 Jan 2023 0 repositories listed
-
Multi-Armed Bandits and Quantum Channel Oracles20 Jan 2023 0 repositories listed
-
Multi-armed Bandit Learning for TDMA Transmission Slot Scheduling and Defragmentation for Improved Bandwidth Usage14 Jan 2023 0 repositories listed
-
Best Arm Identification in Stochastic Bandits: Beyond β-optimality10 Jan 2023 0 repositories listed
-
Local Differential Privacy for Sequential Decision Making in a Changing Environment2 Jan 2023 0 repositories listed
-
Contextual Bandits and Optimistically Universal Learning31 Dec 2022 0 repositories listed
-
Online Statistical Inference for Contextual Bandits via Stochastic Gradient Descent30 Dec 2022 0 repositories listed
-
On the Complexity of Representation Learning in Contextual Linear Bandits19 Dec 2022 0 repositories listed
-
Faster Maximum Inner Product Search in High Dimensions14 Dec 2022 0 repositories listed
-
Corruption-Robust Algorithms with Uncertainty Weighting for Nonlinear Contextual Bandits and Markov Decision Processes12 Dec 2022 0 repositories listed
-
On Regret-optimal Cooperative Nonstochastic Multi-armed Bandits30 Nov 2022 0 repositories listed
-
Constrained Pure Exploration Multi-Armed Bandits with a Fixed Budget27 Nov 2022 0 repositories listed
-
Contextual Decision-Making with Knapsacks Beyond the Worst Case25 Nov 2022 0 repositories listed
-
Contextual Bandits in a Survey Experiment on Charitable Giving: Within-Experiment Outcomes versus Policy Learning22 Nov 2022 0 repositories listed
-
Transfer Learning for Contextual Multi-armed Bandits22 Nov 2022 0 repositories listed
-
Bandit Algorithms for Prophet Inequality and Pandora's Box16 Nov 2022 0 repositories listed
-
Causal Bandits: Online Decision-Making in Endogenous Settings16 Nov 2022 0 repositories listed
-
Multi-Player Bandits Robust to Adversarial Collisions15 Nov 2022 0 repositories listed
-
On Penalization in Stochastic Multi-armed Bandits15 Nov 2022 0 repositories listed
-
Contextual Bandits with Packing and Covering Constraints: A Modular Lagrangian Approach via Regression14 Nov 2022 0 repositories listed
-
Hypothesis Transfer in Bandits by Weighted Models14 Nov 2022 0 repositories listed
-
Generalizing distribution of partial rewards for multi-armed bandits with temporally-partitioned rewards13 Nov 2022 0 repositories listed
-
Contexts can be Cheap: Solving Stochastic Contextual Bandits with Linear Bandit Algorithms8 Nov 2022 0 repositories listed
-
Revisiting Simple Regret: Fast Rates for Returning a Good Arm30 Oct 2022 0 repositories listed
-
Robust Contextual Linear Bandits26 Oct 2022 0 repositories listed
-
PAC-Bayesian Offline Contextual Bandits With Guarantees24 Oct 2022 0 repositories listed
-
Scalable Representation Learning in Linear Contextual Bandits with Constant Regret Guarantees24 Oct 2022 0 repositories listed
-
Vertical Federated Linear Contextual Bandits20 Oct 2022 0 repositories listed
-
Contextual bandits with concave rewards, and an application to fair ranking18 Oct 2022 0 repositories listed
-
Maximum entropy exploration in contextual bandits with neural networks and energy based models12 Oct 2022 0 repositories listed
-
Constant regret for sequence prediction with limited advice5 Oct 2022 0 repositories listed
-
ProtoBandit: Efficient Prototype Selection via Multi-Armed Bandits4 Oct 2022 0 repositories listed
-
Improved High-Probability Regret for Adversarial Bandits with Time-Varying Feedback Graphs4 Oct 2022 0 repositories listed
-
Replicable Bandits4 Oct 2022 0 repositories listed
-
On Best-Arm Identification with a Fixed Budget in Non-Parametric Multi-Armed Bandits30 Sep 2022 0 repositories listed
-
Off-Policy Risk Assessment in Markov Decision Processes21 Sep 2022 0 repositories listed
-
Active Inference for Autonomous Decision-Making with Contextual Multi-Armed Bandits19 Sep 2022 0 repositories listed
-
Towards Robust Off-Policy Evaluation via Human Inputs18 Sep 2022 0 repositories listed
-
Constrained Policy Optimization for Controlled Self-Learning in Conversational AI Systems17 Sep 2022 0 repositories listed
-
Double Doubly Robust Thompson Sampling for Generalized Linear Contextual Bandits15 Sep 2022 0 repositories listed
-
Risk-aware linear bandits with convex loss15 Sep 2022 0 repositories listed
-
Risk-Averse Multi-Armed Bandits with Unobserved Confounders: A Case Study in Emotion Regulation in Mobile Health9 Sep 2022 0 repositories listed
-
Multi-Armed Bandits with Self-Information Rewards6 Sep 2022 0 repositories listed
-
When Privacy Meets Partial Information: A Refined Analysis of Differentially Private Bandits6 Sep 2022 0 repositories listed
-
Exposure-Aware Recommendation using Contextual Bandits4 Sep 2022 0 repositories listed
-
Variational Inference for Model-Free and Model-Based Reinforcement Learning4 Sep 2022 0 repositories listed
-
Dynamic Global Sensitivity for Differentially Private Contextual Bandits30 Aug 2022 0 repositories listed
-
A Provably Efficient Model-Free Posterior Sampling Method for Episodic Reinforcement Learning23 Aug 2022 0 repositories listed
-
Understanding the stochastic dynamics of sequential decision-making processes: A path-integral analysis of multi-armed bandits11 Aug 2022 0 repositories listed
-
Increasing Students' Engagement to Reminder Emails Through Multi-Armed Bandits10 Aug 2022 0 repositories listed
-
Raising Student Completion Rates with Adaptive Curriculum and Contextual Bandits28 Jul 2022 0 repositories listed
-
Towards Soft Fairness in Restless Multi-Armed Bandits27 Jul 2022 0 repositories listed
-
SPRT-based Efficient Best Arm Identification in Stochastic Bandits22 Jul 2022 0 repositories listed
-
Online Learning with Off-Policy Feedback18 Jul 2022 0 repositories listed
-
Parallel Best Arm Identification in Heterogeneous Environments16 Jul 2022 0 repositories listed
-
Model Selection in Reinforcement Learning with General Function Approximations6 Jul 2022 0 repositories listed
-
Instance-optimal PAC Algorithms for Contextual Bandits5 Jul 2022 0 repositories listed
-
Autonomous Drug Design with Multi-Armed Bandits4 Jul 2022 0 repositories listed
-
Joint Representation Training in Sequential Tasks with Shared Structure24 Jun 2022 0 repositories listed
-
Multiple-Play Stochastic Bandits with Shareable Finite-Capacity Arms17 Jun 2022 0 repositories listed
-
A Contextual Combinatorial Semi-Bandit Approach to Network Bottleneck Identification16 Jun 2022 0 repositories listed
-
Combinatorial Pure Exploration of Causal Bandits16 Jun 2022 0 repositories listed
-
Distributed Differential Privacy in Multi-Armed Bandits12 Jun 2022 0 repositories listed
-
Squeeze All: Novel Estimator and Self-Normalized Bound for Linear Contextual Bandits11 Jun 2022 0 repositories listed
-
Communication Efficient Distributed Learning for Kernelized Contextual Bandits10 Jun 2022 0 repositories listed
-
Conformal Off-Policy Prediction in Contextual Bandits9 Jun 2022 0 repositories listed
-
Efficient Resource Allocation with Fairness Constraints in Restless Multi-Armed Bandits8 Jun 2022 0 repositories listed
-
Neural Bandit with Arm Group Graph8 Jun 2022 0 repositories listed
-
A Simple and Optimal Policy Design with Safety against Heavy-Tailed Risk for Stochastic Bandits7 Jun 2022 0 repositories listed
-
Finite-Time Regret of Thompson Sampling Algorithms for Exponential Family Multi-Armed Bandits7 Jun 2022 0 repositories listed
-
Asymptotic Instance-Optimal Algorithms for Interactive Decision Making6 Jun 2022 0 repositories listed
-
Robust Pareto Set Identification with Contaminated Bandit Feedback6 Jun 2022 0 repositories listed
-
Contextual Bandits with Knapsacks for a Conversion Model1 Jun 2022 0 repositories listed
-
Online Meta-Learning in Adversarial Multi-Armed Bandits31 May 2022 0 repositories listed
-
Provable General Function Class Representation Learning in Multitask Bandits and MDPs31 May 2022 0 repositories listed
-
Provably and Practically Efficient Neural Contextual Bandits31 May 2022 0 repositories listed
-
Quantum Multi-Armed Bandits and Stochastic Linear Bandits Enjoy Logarithmic Regrets30 May 2022 0 repositories listed
-
Fairness and Welfare Quantification for Regret in Multi-Armed Bandits27 May 2022 0 repositories listed
-
Lifting the Information Ratio: An Information-Theoretic Analysis of Thompson Sampling for Contextual Bandits27 May 2022 0 repositories listed
-
Meta-Learning Adversarial Bandits27 May 2022 0 repositories listed
-
Contextual Pandora's Box26 May 2022 0 repositories listed
-
Exploration, Exploitation, and Engagement in Multi-Armed Bandits with Abandonment26 May 2022 0 repositories listed
-
Neural Contextual Bandits Based Dynamic Sensor Selection for Low-Power Body-Area Networks24 May 2022 0 repositories listed
-
Computationally Efficient Horizon-Free Reinforcement Learning for Linear Mixture MDPs23 May 2022 0 repositories listed
-
Falsification of Multiple Requirements for Cyber-Physical Systems Using Online Generative Adversarial Networks and Multi-Armed Bandits23 May 2022 0 repositories listed
-
Contextual Information-Directed Sampling22 May 2022 0 repositories listed
-
Pessimism for Offline Linear Contextual Bandits using ℓₚ Confidence Sets21 May 2022 0 repositories listed
-
Stability Enforced Bandit Algorithms for Channel Selection in Remote State Estimation of Gauss-Markov Processes20 May 2022 0 repositories listed