Browse State-of-the-Art › Sequential Decision Making › Papers, page 4
Sequential Decision Making
Papers archive 2025-07-28
archive papers tagged: 1,210 · with a code link: 351 · where Syntology ran a sample: 107 (90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument) Syntology
Show: all tagged papersonly where code ran (107 of 1,210 tagged: 90 with a run with no instrument failure, 17 where every run was a failure of Syntology's instrument)
Page 4 of 13: papers 301 to 400 of 1,210, in archive order: by repositories listed in the archive (most first), then newest first, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers that list no repository come after every paper that lists one.
Papers without a page here are shown as plain text. A Syntology line reads “N ran (of which C constructed an object rather than computing a result; K with no instrument failure: H honoured, V violated, P with no contract checked; I where Syntology's instrument failed) · U unverified”; the figure “where Syntology's instrument failed” counts failures of Syntology's instrument, not of the code. When the archive marks a repository official for the paper, the line starts with that repository's state (the archive's flag, not a verdict on who wrote the code); hover it for the repositories the samples that ran came from. Abstracts are on each paper's page.
-
13 Feb 2020 1 repository listed
-
27 Jan 2020 1 repository listed
-
13 Jan 2020 1 repository listed Syntology official (archive's flag): 2 ran · 2 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 2 where Syntology's instrument failed) · 0 unverified (of 2 harvested samples)
-
3 Dec 2019 1 repository listed
-
1 Dec 2019 1 repository listed
-
19 Nov 2019 1 repository listed Syntology 4 ran (of which 0 constructed an object rather than computing a result; 4 with no instrument failure: 0 honoured, 0 violated, 4 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 5 harvested samples) · 3 pointer-only (licence)
-
11 Nov 2019 1 repository listed
-
2 Nov 2019 1 repository listed
-
30 Oct 2019 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
15 Oct 2019 1 repository listed
-
4 Oct 2019 1 repository listed
-
4 Oct 2019 1 repository listed
-
11 Sep 2019 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples)
-
8 Sep 2019 1 repository listed
-
27 Aug 2019 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 5 with no instrument failure: 0 honoured, 0 violated, 5 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 6 harvested samples) · 6 pointer-only (licence)
-
30 Jul 2019 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample)
-
25 Jul 2019 1 repository listed
-
3 Jul 2019 1 repository listed
-
1 Jul 2019 1 repository listed
-
20 Jun 2019 1 repository listed
-
5 Jun 2019 1 repository listed
-
5 Jun 2019 1 repository listed
-
27 May 2019 1 repository listed
-
23 May 2019 1 repository listed
-
22 Apr 2019 1 repository listed
-
15 Feb 2019 1 repository listed
-
5 Feb 2019 1 repository listed
-
24 Jan 2019 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 1 honoured, 0 violated, 0 with no contract checked; 0 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
21 Jan 2019 1 repository listed
-
1 Dec 2018 1 repository listed
-
29 Oct 2018 1 repository listed
-
30 Sep 2018 1 repository listed
-
8 Aug 2018 1 repository listed
-
17 Jul 2018 1 repository listed Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 0 honoured, 0 violated, 6 with no contract checked; 1 where Syntology's instrument failed) · 3 unverified (of 10 harvested samples) · 1 pointer-only (licence)
-
21 Jun 2018 1 repository listed
-
6 Jun 2018 1 repository listed Syntology official (archive's flag): 7 ran · 7 ran (of which 0 constructed an object rather than computing a result; 7 with no instrument failure: 0 honoured, 0 violated, 7 with no contract checked; 0 where Syntology's instrument failed) · 2 unverified (of 9 harvested samples)
-
25 May 2018 1 repository listed
-
20 May 2018 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 0 with no instrument failure: 0 honoured, 0 violated, 0 with no contract checked; 1 where Syntology's instrument failed) · 0 unverified (of 1 harvested sample) · 1 pointer-only (licence)
-
17 Apr 2018 1 repository listed Syntology official (archive's flag): 6 ran · 6 ran (of which 0 constructed an object rather than computing a result; 6 with no instrument failure: 1 honoured, 0 violated, 5 with no contract checked; 0 where Syntology's instrument failed) · 4 unverified (of 10 harvested samples) · 10 pointer-only (licence)
-
22 Feb 2018 1 repository listed
-
6 Feb 2018 1 repository listed
-
30 Dec 2017 1 repository listed
-
20 Nov 2017 1 repository listed
-
26 Apr 2017 1 repository listed
-
6 Feb 2016 1 repository listed
-
10 Jun 2015 1 repository listed Syntology official (archive's flag): 1 ran · 1 ran (of which 0 constructed an object rather than computing a result; 1 with no instrument failure: 0 honoured, 0 violated, 1 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 2 harvested samples)
-
10 Mar 2015 1 repository listed
-
1 Dec 2009 1 repository listed
-
AirLLM: Diffusion Policy-based Adaptive LoRA for Remote Fine-Tuning of LLM over the Air15 Jul 2025 0 repositories listed
-
LLM-Stackelberg Games: Conjectural Reasoning Equilibria and Their Applications to Spearphishing12 Jul 2025 0 repositories listed
-
A Survey of Continual Reinforcement Learning27 Jun 2025 0 repositories listed
-
26 Jun 2025 0 repositories listed Syntology 3 ran (of which 2 constructed an object rather than computing a result; 3 with no instrument failure: 1 honoured, 0 violated, 2 with no contract checked; 0 where Syntology's instrument failed) · 1 unverified (of 4 harvested samples) · 4 pointer-only (licence)
-
POLAR: A Pessimistic Model-based Policy Learning Algorithm for Dynamic Treatment Regimes25 Jun 2025 0 repositories listed
-
Efficient Strategy Synthesis for MDPs via Hierarchical Block Decomposition21 Jun 2025 0 repositories listed
-
Multi-Armed Bandits With Machine Learning-Generated Surrogate Rewards20 Jun 2025 0 repositories listed
-
Leveraging In-Context Learning for Language Model Agents16 Jun 2025 0 repositories listed
-
Revisiting Clustering of Neural Bandits: Selective Reinitialization for Mitigating Loss of Plasticity14 Jun 2025 0 repositories listed
-
TooBadRL: Trigger Optimization to Boost Effectiveness of Backdoor Attacks on Deep Reinforcement Learning11 Jun 2025 0 repositories listed
-
Towards Responsible AI: Advances in Safety, Fairness, and Accountability of Autonomous Systems11 Jun 2025 0 repositories listed
-
How to Provably Improve Return Conditioned Supervised Learning?10 Jun 2025 0 repositories listed
-
QForce-RL: Quantized FPGA-Optimized Reinforcement Learning Compute Engine8 Jun 2025 0 repositories listed
-
Contextual Experience Replay for Self-Improvement of Language Agents7 Jun 2025 0 repositories listed
-
AutoQD: Automatic Discovery of Diverse Behaviors with Quality-Diversity Optimization5 Jun 2025 0 repositories listed
-
Active Layer-Contrastive Decoding Reduces Hallucination in Large Language Model Generation29 May 2025 0 repositories listed
-
Emergent Risk Awareness in Rational Agents under Resource Constraints29 May 2025 0 repositories listed
-
Adaptive Frontier Exploration on Graphs with Applications to Network-Based Disease Testing27 May 2025 0 repositories listed
-
Variational Deep Learning via Implicit Regularization26 May 2025 0 repositories listed
-
DDO: Dual-Decision Optimization via Multi-Agent Collaboration for LLM-Based Medical Consultation24 May 2025 0 repositories listed
-
Automata Learning of Preferences over Temporal Logic Formulas from Pairwise Comparisons23 May 2025 0 repositories listed
-
Reward Is Enough: LLMs Are In-Context Reinforcement Learners21 May 2025 0 repositories listed
-
Vid2World: Crafting Video Diffusion Models to Interactive World Models20 May 2025 0 repositories listed
-
OMGPT: A Sequence Modeling Framework for Data-driven Operational Decision Making19 May 2025 0 repositories listed
-
Deep Symbolic Optimization: Reinforcement Learning for Symbolic Mathematics16 May 2025 0 repositories listed
-
Generalization Guarantees for Learning Branch-and-Cut Policies in Integer Programming16 May 2025 0 repositories listed
-
Batched Nonparametric Bandits via k-Nearest Neighbor UCB15 May 2025 0 repositories listed
-
Counterfactual Strategies for Markov Decision Processes14 May 2025 0 repositories listed
-
Sequential Treatment Effect Estimation with Unmeasured Confounders14 May 2025 0 repositories listed
-
rfPG: Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs14 May 2025 0 repositories listed
-
A Practical Introduction to Deep Reinforcement Learning13 May 2025 0 repositories listed
-
Explainable Reinforcement Learning Agents Using World Models12 May 2025 0 repositories listed
-
A Multi-Agent Reinforcement Learning Approach for Cooperative Air-Ground-Human Crowdsensing in Emergency Rescue11 May 2025 0 repositories listed
-
Constrained Online Decision-Making: A Unified Framework11 May 2025 0 repositories listed
-
RL-DAUNCE: Reinforcement Learning-Driven Data Assimilation with Uncertainty-Aware Constrained Ensembles8 May 2025 0 repositories listed
-
MDPs with a State Sensing Cost6 May 2025 0 repositories listed
-
Policy-labeled Preference Learning: Is Preference Enough for RLHF?6 May 2025 0 repositories listed
-
D3HRL: A Distributed Hierarchical Reinforcement Learning Approach Based on Causal Discovery and Spurious Correlation Detection4 May 2025 0 repositories listed
-
Bayesian learning of the optimal action-value function in a Markov decision process3 May 2025 0 repositories listed
-
A Minimax-MDP Framework with Future-imposed Conditions for Learning-augmented Problems2 May 2025 0 repositories listed
-
Self-Generated In-Context Examples Improve LLM Agents for Sequential Decision-Making Tasks1 May 2025 0 repositories listed
-
Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments27 Apr 2025 0 repositories listed
-
SAPO-RL: Sequential Actuator Placement Optimization for Fuselage Assembly via Reinforcement Learning24 Apr 2025 0 repositories listed
-
Hierarchical Attention Fusion of Visual and Textual Representations for Cross-Domain Sequential Recommendation21 Apr 2025 0 repositories listed
-
Consensus in Motion: A Case of Dynamic Rationality of Sequential Learning in Probability Aggregation20 Apr 2025 0 repositories listed
-
TALES: Text Adventure Learning Environment Suite19 Apr 2025 0 repositories listed
-
Position Paper: Rethinking Privacy in RL for Sequential Decision-making in the Age of LLMs15 Apr 2025 0 repositories listed
-
Truncated Matrix Completion - An Empirical Study14 Apr 2025 0 repositories listed
-
Towards More Efficient, Robust, Instance-adaptive, and Generalizable Sequential Decision making12 Apr 2025 0 repositories listed
Syntology lines on 16 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced; each line links to that paper's sample list. Syntology's record for this page has not changed since , the first build that kept a record date for it; when this build read Syntology's graph is in the build record. For agents: get_harvested_code_for_paper(arxiv_id) lists each paper's samples; how to connect.