Browse State-of-the-Art › Thompson Sampling › Papers, page 5
Thompson Sampling
Papers archive 2025-07-28
archive papers tagged: 655 · with a code link: 135 · where Syntology ran a sample: 30 (23 with a run with no instrument failure, 7 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (30 of 655 tagged: 23 with a run with no instrument failure, 7 where every run was a failure of Syntology's instrument)
Page 5 of 7: papers 401 to 500 of 655, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
Asymptotically Optimal Bandits under Weighted Information28 May 2021 0 repositories listed
-
Diffusion Approximations for Thompson Sampling19 May 2021 0 repositories listed
-
Thompson Sampling for Gaussian Entropic Risk Bandits14 May 2021 0 repositories listed
-
High-dimensional near-optimal experiment design for drug discovery via Bayesian sparse sampling23 Apr 2021 0 repositories listed
-
When and Whom to Collaborate with in a Changing Environment: A Collaborative Dynamic Bandit Solution14 Apr 2021 0 repositories listed
-
Blind Exploration and Exploitation of Stochastic Experts2 Apr 2021 0 repositories listed
-
Challenges in Statistical Analysis of Data Collected by a Bandit Algorithm: An Empirical Exploration in Applications to Adaptively Randomized Experiments22 Mar 2021 0 repositories listed
-
Constrained Contextual Bandit Learning for Adaptive Radar Waveform Selection9 Mar 2021 0 repositories listed
-
Online Multi-Armed Bandits with Adaptive Inference25 Feb 2021 0 repositories listed
-
Model-based Meta Reinforcement Learning using Graph Structured Surrogate Models16 Feb 2021 0 repositories listed
-
Near-Optimal Algorithms for Differentially Private Online Learning in a Stochastic Environment16 Feb 2021 0 repositories listed
-
The Elliptical Potential Lemma for General Distributions with an Application to Linear Thompson Sampling16 Feb 2021 0 repositories listed
-
Meta-Thompson Sampling11 Feb 2021 0 repositories listed
-
Doubly robust Thompson sampling for linear payoffs1 Feb 2021 0 repositories listed
-
Weak Signal Asymptotics for Sequentially Randomized Experiments25 Jan 2021 0 repositories listed
-
TSEC: a framework for online experimentation under experimental constraints17 Jan 2021 0 repositories listed
-
Deciding What to Learn: A Rate-Distortion Approach15 Jan 2021 0 repositories listed
-
Etat de l'art sur l'application des bandits multi-bras4 Jan 2021 0 repositories listed
-
Meta-Reinforcement Learning With Informed Policy Regularization1 Jan 2021 0 repositories listed
-
Aging Bandits: Regret Analysis and Order-Optimal Learning Algorithm for Wireless Networks with Stochastic Arrivals16 Dec 2020 0 repositories listed
-
Reinforcement Learning with Subspaces using Free Energy Paradigm13 Dec 2020 0 repositories listed
-
Distributed Thompson Sampling3 Dec 2020 0 repositories listed
-
Non-Stationary Latent Bandits1 Dec 2020 0 repositories listed
-
On Efficiency in Hierarchical Reinforcement Learning1 Dec 2020 0 repositories listed
-
Distilled Thompson Sampling: Practical and Efficient Thompson Sampling via Imitation Learning29 Nov 2020 0 repositories listed
-
Reward Biased Maximum Likelihood Estimation for Reinforcement Learning16 Nov 2020 0 repositories listed
-
Risk-Constrained Thompson Sampling for CVaR Bandits16 Nov 2020 0 repositories listed
-
Accelerating Grasp Exploration by Leveraging Learned Priors11 Nov 2020 0 repositories listed
-
Thompson sampling for linear quadratic mean-field teams9 Nov 2020 0 repositories listed
-
Asymptotic Convergence of Thompson Sampling8 Nov 2020 0 repositories listed
-
Adaptive Combinatorial Allocation4 Nov 2020 0 repositories listed
-
Greedy k-Center from Noisy Distance Samples3 Nov 2020 0 repositories listed
-
Screening for an Infectious Disease as a Problem in Stochastic Control1 Nov 2020 0 repositories listed
-
Bandit Policies for Reliable Cellular Network Handovers in Extreme Mobility28 Oct 2020 0 repositories listed
-
Improved Worst-Case Regret Bounds for Randomized Least-Squares Value Iteration23 Oct 2020 0 repositories listed
-
Reinforcement Learning for Efficient and Tuning-Free Link Adaptation16 Oct 2020 0 repositories listed
-
Double-Linear Thompson Sampling for Context-Attentive Bandits15 Oct 2020 0 repositories listed
-
Online Learning and Distributed Control for Residential Demand Response11 Oct 2020 0 repositories listed
-
Effects of Model Misspecification on Bayesian Bandits: Case Studies in UX Optimization7 Oct 2020 0 repositories listed
-
Stage-wise Conservative Linear Bandits30 Sep 2020 0 repositories listed
-
Neural Model-based Optimization with Right-Censored Observations29 Sep 2020 0 repositories listed
-
Position-Based Multiple-Play Bandits with Thompson Sampling28 Sep 2020 0 repositories listed
-
Bandit Change-Point Detection for Real-Time Monitoring High-Dimensional Data Under Sampling Control24 Sep 2020 0 repositories listed
-
Partially Observable Online Change Detection via Smooth-Sparse Decomposition22 Sep 2020 0 repositories listed
-
Bandits Under The Influence (Extended Version)21 Sep 2020 0 repositories listed
-
Causal Bandits without prior knowledge using separating sets16 Sep 2020 0 repositories listed
-
Thompson Sampling for Unsupervised Sequential Selection16 Sep 2020 0 repositories listed
-
A Change-Detection Based Thompson Sampling Framework for Non-Stationary Bandits6 Sep 2020 0 repositories listed
-
Efficient Online Learning for Cognitive Radar-Cellular Coexistence via Contextual Thompson Sampling24 Aug 2020 0 repositories listed
-
Contextual Bandits for Advertising Budget Allocation22 Aug 2020 0 repositories listed
-
Near Optimal Adversarial Attacks on Stochastic Bandits and Defenses with Smoothed Responses21 Aug 2020 0 repositories listed
-
Reinforcement Learning with Trajectory Feedback13 Aug 2020 0 repositories listed
-
Lenient Regret for Multi-Armed Bandits10 Aug 2020 0 repositories listed
-
IntelligentPooling: Practical Thompson Sampling for mHealth31 Jul 2020 0 repositories listed
-
Greedy Bandits with Sampled Context27 Jul 2020 0 repositories listed
-
Influence Diagram Bandits: Variational Thompson Sampling for Structured Bandit Problems9 Jul 2020 0 repositories listed
-
Variable Selection via Thompson Sampling1 Jul 2020 0 repositories listed
-
Policy Gradient Optimization of Thompson Sampling Policies30 Jun 2020 0 repositories listed
-
Asynchronous Multi Agent Active Search25 Jun 2020 0 repositories listed
-
Learning by Repetition: Stochastic Multi-armed Bandits under Priming Effect18 Jun 2020 0 repositories listed
-
Analysis and Design of Thompson Sampling for Stochastic Partial Monitoring17 Jun 2020 0 repositories listed
-
Constrained Thompson Sampling for Real-Time Electricity Pricing with Grid Reliability Constraints17 Jun 2020 0 repositories listed
-
Latent Bandits Revisited15 Jun 2020 0 repositories listed
-
Hypermodels for Exploration12 Jun 2020 0 repositories listed
-
On Frequentist Regret of Linear Thompson Sampling11 Jun 2020 0 repositories listed
-
Statistical Efficiency of Thompson Sampling for Combinatorial Semi-Bandits11 Jun 2020 0 repositories listed
-
TS-UCB: Improving on Thompson Sampling With Little to No Additional Computation11 Jun 2020 0 repositories listed
-
Scalable Thompson Sampling using Sparse Gaussian Process Models9 Jun 2020 0 repositories listed
-
Random Hypervolume Scalarizations for Provable Multi-Objective Black Box Optimization8 Jun 2020 0 repositories listed
-
An Efficient Algorithm For Generalized Linear Bandit: Online Stochastic Gradient Descent and Thompson Sampling7 Jun 2020 0 repositories listed
-
Concurrent Decentralized Channel Allocation and Access Point Selection using Multi-Armed Bandits in multi BSS WLANs5 Jun 2020 0 repositories listed
-
Thompson Sampling for Combinatorial Semi-bandits with Sleeping Arms and Long-Term Fairness Constraints14 May 2020 0 repositories listed
-
Learning to Rank in the Position Based Model with Bandit Feedback27 Apr 2020 0 repositories listed
-
Online Learning with Cumulative Oversampling: Application to Budgeted Influence Maximization24 Apr 2020 0 repositories listed
-
Adaptive Operator Selection Based on Dynamic Thompson Sampling for MOEA/D22 Apr 2020 0 repositories listed
-
Optimal No-regret Learning in Repeated First-price Auctions22 Mar 2020 0 repositories listed
-
A Reliability-aware Multi-armed Bandit Approach to Learn and Select Users in Demand Response20 Mar 2020 0 repositories listed
-
Delay-Adaptive Learning in Generalized Linear Contextual Bandits11 Mar 2020 0 repositories listed
-
Online Residential Demand Response via Contextual Multi-Armed Bandits7 Mar 2020 0 repositories listed
-
An Online Learning Framework for Energy-Efficient Navigation of Electric Vehicles3 Mar 2020 0 repositories listed
-
MOTS: Minimax Optimal Thompson Sampling3 Mar 2020 0 repositories listed
-
Efficient exploration of zero-sum stochastic games24 Feb 2020 0 repositories listed
-
On Thompson Sampling with Langevin Algorithms23 Feb 2020 0 repositories listed
-
Residual Bootstrap Exploration for Bandit Algorithms19 Feb 2020 0 repositories listed
-
A General Theory of the Stochastic Linear Bandit and Its Applications12 Feb 2020 0 repositories listed
-
The Price of Incentivizing Exploration: A Characterization via Thompson Sampling and Sample Complexity3 Feb 2020 0 repositories listed
-
Bayesian Quantile and Expectile Optimisation12 Jan 2020 0 repositories listed
-
On Thompson Sampling for Smoother-than-Lipschitz Bandits8 Jan 2020 0 repositories listed
-
Solving Bernoulli Rank-One Bandits with Unimodal Thompson Sampling6 Dec 2019 0 repositories listed
-
Ordinal Bayesian Optimisation5 Dec 2019 0 repositories listed
-
Thompson Sampling and Approximate Inference1 Dec 2019 0 repositories listed
-
Automatic Ensemble Learning for Online Influence Maximization25 Nov 2019 0 repositories listed
-
Information-Theoretic Confidence Bounds for Reinforcement Learning21 Nov 2019 0 repositories listed
-
Adaptive Portfolio by Solving Multi-armed Bandit via Thompson Sampling13 Nov 2019 0 repositories listed
-
Incentivized Exploration for Multi-Armed Bandits under Reward Drift12 Nov 2019 0 repositories listed
-
Safe Linear Thompson Sampling with Side Information6 Nov 2019 0 repositories listed
-
On Batch Bayesian Optimization4 Nov 2019 0 repositories listed
-
On Online Learning in Kernelized Markov Decision Processes4 Nov 2019 0 repositories listed
-
Fixed-Confidence Guarantees for Bayesian Best-Arm Identification24 Oct 2019 0 repositories listed
Syntology lines on 1 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.