Browse State-of-the-Art › Thompson Sampling › Papers, page 2
Thompson Sampling
Papers archive 2025-07-28
archive papers tagged: 655 · with a code link: 135 · where Syntology ran a sample: 30 (23 with a run with no instrument failure, 7 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (30 of 655 tagged: 23 with a run with no instrument failure, 7 where every run was a failure of Syntology's instrument)
Page 2 of 7: papers 101 to 200 of 655, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
4 Mar 2020 1 repository listed
-
27 Feb 2020 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 2 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
1 Dec 2019 1 repository listed
-
28 Nov 2019 1 repository listed
-
22 Nov 2019 1 repository listed
-
2 Nov 2019 1 repository listed
-
30 Oct 2019 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
11 Oct 2019 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 3 with no instrument failure: 3 honoured, 0 violated, 0 with no contract checked; 3 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
25 Jul 2019 1 repository listed
-
11 Jul 2019 1 repository listed
-
29 May 2019 1 repository listed
-
10 May 2019 1 repository listed
-
11 Mar 2019 1 repository listed
-
12 Feb 2019 1 repository listed
-
23 Jan 2019 1 repository listed
-
18 Dec 2018 1 repository listed
-
11 Dec 2018 1 repository listed
-
1 Dec 2018 1 repository listed
-
8 Nov 2018 1 repository listed
-
8 Aug 2018 1 repository listed
-
8 Aug 2018 1 repository listed
-
25 May 2018 1 repository listed
-
13 Feb 2018 1 repository listed
-
10 Sep 2017 1 repository listed
-
10 Sep 2017 1 repository listed
-
25 May 2017 1 repository listed
-
22 May 2017 1 repository listed
-
28 Apr 2017 1 repository listed
-
28 Feb 2017 1 repository listed
-
25 Apr 2016 1 repository listed
-
17 Mar 2016 1 repository listed
-
26 Feb 2016 1 repository listed
-
3 Jul 2015 1 repository listed
-
2 Jun 2015 1 repository listed
-
18 May 2012 1 repository listed
-
Robust Policy Switching for Antifragile Reinforcement Learning for UAV Deconfliction in Adversarial Environments26 Jun 2025 0 repositories listed
-
Context Attribution with Multi-Armed Bandit Optimization24 Jun 2025 0 repositories listed
-
Adaptive Data Augmentation for Thompson Sampling17 Jun 2025 0 repositories listed
-
Bayesian Optimization with Inexact Acquisition: Is Random Grid Search Sufficient?13 Jun 2025 0 repositories listed
-
Efficient kernelized bandit algorithms via exploration distributions11 Jun 2025 0 repositories listed
-
Asymptotically Optimal Linear Best Feasible Arm Identification with Fixed Budget3 Jun 2025 0 repositories listed
-
Simplifying Bayesian Optimization Via In-Context Direct Optimum Sampling29 May 2025 0 repositories listed
-
Stable Thompson Sampling: Valid Inference via Variance Inflation29 May 2025 0 repositories listed
-
Thompson Sampling in Online RLHF with General Function Approximation29 May 2025 0 repositories listed
-
Practical Adversarial Attacks on Stochastic Bandits via Fake Data Injection28 May 2025 0 repositories listed
-
Deconfounded Warm-Start Thompson Sampling with Applications to Precision Medicine22 May 2025 0 repositories listed
-
Generator-Mediated Bandits: Thompson Sampling for GenAI-Powered Adaptive Interventions22 May 2025 0 repositories listed
-
Scalable and Interpretable Contextual Bandits: A Literature Review and Retail Offer Prototype22 May 2025 0 repositories listed
-
In-Domain African Languages Translation Using LLMs and Multi-armed Bandits21 May 2025 0 repositories listed
-
Dynamic Decision-Making under Model Misspecification20 May 2025 0 repositories listed
-
Thompson Sampling-like Algorithms for Stochastic Rising Bandits17 May 2025 0 repositories listed
-
Leveraging Offline Data from Similar Systems for Online Linear Quadratic Control14 May 2025 0 repositories listed
-
Connecting Thompson Sampling and UCB: Towards More Efficient Trade-offs Between Privacy and Regret5 May 2025 0 repositories listed
-
Bayesian learning of the optimal action-value function in a Markov decision process3 May 2025 0 repositories listed
-
Neural Contextual Bandits Under Delayed Feedback Constraints16 Apr 2025 0 repositories listed
-
Counterfactual Inference under Thompson Sampling3 Apr 2025 0 repositories listed
-
Sparse Nonparametric Contextual Bandits20 Mar 2025 0 repositories listed
-
Achieving adaptivity and optimality for multi-armed bandits using Exponential-Kullback Leibler Maillard Sampling20 Feb 2025 0 repositories listed
-
An Adversarial Analysis of Thompson Sampling for Full-information Online Learning: from Finite to Infinite Action Spaces20 Feb 2025 0 repositories listed
-
Uncertainty-Aware Search and Value Models: Mitigating Search Scaling Flaws in LLMs16 Feb 2025 0 repositories listed
-
When and why randomised exploration works (in linear bandits)13 Feb 2025 0 repositories listed
-
KABB: Knowledge-Aware Bayesian Bandits for Dynamic Expert Coordination in Multi-Agent Systems11 Feb 2025 0 repositories listed
-
Contextual Thompson Sampling via Generation of Missing Data10 Feb 2025 0 repositories listed
-
An Information-Theoretic Analysis of Thompson Sampling with Infinite Action Spaces4 Feb 2025 0 repositories listed
-
Active RLHF via Best Policy Learning from Trajectory Preference Feedback31 Jan 2025 0 repositories listed
-
31 Jan 2025 0 repositories listed Syntology 2 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 1 unverified (of 3 harvested samples) · 3 pointer-only (licence)
-
EVaDE : Event-Based Variational Thompson Sampling for Model-Based Reinforcement Learning16 Jan 2025 0 repositories listed
-
Stochastically Constrained Best Arm Identification with Thompson Sampling7 Jan 2025 0 repositories listed
-
Truthful mechanisms for linear bandit games with private contexts7 Jan 2025 0 repositories listed
-
WAPTS: A Weighted Allocation Probability Adjusted Thompson Sampling Algorithm for High-Dimensional and Sparse Experiment Settings7 Jan 2025 0 repositories listed
-
On Improved Regret Bounds In Bayesian Optimization with Gaussian Noise25 Dec 2024 0 repositories listed
-
Generalized Bayesian deep reinforcement learning16 Dec 2024 0 repositories listed
-
An Information-Theoretic Analysis of Thompson Sampling for Logistic Bandits3 Dec 2024 0 repositories listed
-
30 Nov 2024 0 repositories listed Syntology 3 ran (of which 2 constructed an object rather than computing a result; 2 with no instrument failure: 0 honoured, 0 violated, 2 with no contract checked; 1 where Syntology's instrument failed) · 6 unverified (of 9 harvested samples) · 9 pointer-only (licence)
-
Epinet for Content Cold Start20 Nov 2024 0 repositories listed
-
Planning and Learning in Risk-Aware Restless Multi-Arm Bandit Problem30 Oct 2024 0 repositories listed
-
BanditCAT and AutoIRT: Machine Learning Approaches to Computerized Adaptive Testing and Item Calibration28 Oct 2024 0 repositories listed
-
Bayesian Collaborative Bandits with Thompson Sampling for Improved Outreach in Maternal Health Program28 Oct 2024 0 repositories listed
-
Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks25 Oct 2024 0 repositories listed
-
Aligning AI Agents via Information-Directed Sampling18 Oct 2024 0 repositories listed
-
Combinatorial Multi-armed Bandits: Arm Selection via Group Testing14 Oct 2024 0 repositories listed
-
Gaussian Process Thompson Sampling via Rootfinding10 Oct 2024 0 repositories listed
-
Contextual Bandits with Non-Stationary Correlated Rewards for User Association in MmWave Vehicular Networks8 Oct 2024 0 repositories listed
-
Efficient Model-Based Reinforcement Learning Through Optimistic Thompson Sampling7 Oct 2024 0 repositories listed
-
Partially Observable Contextual Bandits with Linear Payoffs17 Sep 2024 0 repositories listed
-
Modified Meta-Thompson Sampling for Linear Bandits and Its Bayes Regret Analysis10 Sep 2024 0 repositories listed
-
Sliding-Window Thompson Sampling for Non-Stationary Settings8 Sep 2024 0 repositories listed
-
Multi-Task Combinatorial Bandits for Budget Allocation31 Aug 2024 0 repositories listed
-
An Extremely Data-efficient and Generative LLM-based Reinforcement Learning Agent for Recommenders28 Aug 2024 0 repositories listed
-
Improving Thompson Sampling via Information Relaxation for Budgeted Multi-armed Bandits28 Aug 2024 0 repositories listed
-
Contextual Bandit with Herding Effects: Algorithms and Recommendation Applications26 Aug 2024 0 repositories listed
-
Optimization-Driven Adaptive Experimentation8 Aug 2024 0 repositories listed
-
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback24 Jul 2024 0 repositories listed
-
Thompson Sampling Itself is Differentially Private20 Jul 2024 0 repositories listed
-
DRL-based Joint Resource Scheduling of eMBB and URLLC in O-RAN16 Jul 2024 0 repositories listed
-
Bayesian Bandit Algorithms with Approximate Inference in Stochastic Linear Bandits20 Jun 2024 0 repositories listed
-
Joint User Association and Pairing in Multi-UAV-Assisted NOMA Networks: A Decaying-Epsilon Thompson Sampling Framework20 Jun 2024 0 repositories listed
-
Preferential Multi-Objective Bayesian Optimization20 Jun 2024 0 repositories listed
-
Memory Sequence Length of Data Sampling Impacts the Adaptation of Meta-Reinforcement Learning Agents18 Jun 2024 0 repositories listed
-
Improving Reward-Conditioned Policies for Multi-Armed Bandits using Normalized Weight Functions16 Jun 2024 0 repositories listed
Syntology lines on 5 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.