Browse State-of-the-Art › Sequential Decision Making
Sequential Decision Making
351 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 351 papers with code (1,210 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
29 Dec 2017 6 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedVideo summarization aims to facilitate large-scale video browsing by producing short, concise summaries that are diverse and representative of original videos.
-
20 Mar 2023 5 repositories listed Syntology ran 2 of 9 samples · 7 unverifiedLarge language models (LLMs) have been increasingly used to interact with external environments (e.
-
23 Jun 2021 5 repositories listed Syntology ran 2 of 3 samples · 1 unverified · 3 pointer-only (licence)In many sequential decision-making problems (e.
-
29 Oct 2018 5 repositories listedThe DRR framework treats recommendation as a sequential decision making procedure and adopts an "Actor-Critic" reinforcement learning scheme to model the interactions between the users and recommender systems, which can…
-
26 Feb 2018 4 repositories listedAt the same time, advances in approximate Bayesian methods have made posterior approximation for flexible neural network models practical.
-
4 Dec 2017 4 repositories listedHierarchical agents have the potential to solve sequential decision making tasks with greater sample efficiency than their non-hierarchical counterparts because hierarchical agents can break down tasks into sets of…
-
23 May 2017 4 repositories listedSequential decision making problems, such as structured prediction, robotic control, and game playing, require a combination of planning policies and generalisation of those plans.
-
3 Oct 2022 3 repositories listedTo help answer this, we first introduce an open-source modular library, RL4LMs (Reinforcement Learning for Language Models), for optimizing language generators with RL.
-
5 Jan 2019 3 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedThe agent finally finds an optimal classification policy in imbalanced data under the guidance of specific reward function and beneficial learning environment.
-
14 Sep 2017 3 repositories listedImage cropping aims at improving the aesthetic quality of images by adjusting their composition.
-
16 Dec 2016 3 repositories listedA softmax operator applied to a set of values acts somewhat like the maximization function and somewhat like an average.
-
14 Jun 2016 3 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedState of the art deep reinforcement learning algorithms take many millions of interactions to attain human-level performance.
-
19 Aug 2024 2 repositories listedReinforcement Learning (RL) is a potent tool for sequential decision-making and has achieved performance surpassing human capabilities across many challenging real-world tasks.
-
18 Jul 2024 2 repositories listed Syntology ran 5 of 8 samples · 3 unverified · 8 pointer-only (licence)Thompson Sampling is a principled method for balancing exploration and exploitation, but its real-world adoption faces computational challenges in large-scale or non-conjugate settings.
-
7 Feb 2024 2 repositories listedOur results demonstrate that Sym-Q excels not only in recovering underlying mathematical structures but also uniquely learns to efficiently refine the output expression based on reward signals, thereby discovering…
-
30 Jan 2024 2 repositories listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)Although reinforcement learning (RL) can solve many challenging sequential decision making problems, achieving zero-shot transfer across related tasks remains a challenge.
-
7 Sep 2023 2 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedTo address biases in sequential decision-making, we introduce a long-term fairness concept named Equal Long-term Benefit Rate (ELBERT).
-
5 Jun 2023 2 repositories listedWe present an information-theoretic framework to learn fixed-dimensional embeddings for tasks in reinforcement learning.
-
16 Feb 2023 2 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedWe present Trieste, an open-source Python package for Bayesian optimization and active learning benefiting from the scalability and efficiency of TensorFlow.
-
30 May 2022 2 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedA slow stream that is recurrent in nature aims to learn a specialized and compressed representation, by forcing chunks of K time steps into a single representation which is divided into multiple vectors.
-
1 Dec 2021 2 repositories listedTo address this issue, we devise an adaptive mechanism to align reinforcement learning and classification methods using distribution entropy as the medium.
-
23 Feb 2021 2 repositories listed Syntology ran 1 of 1 samples · 0 unverified · 1 pointer-only (licence)Formally, MPG is constructed as a weighted average of the data-driven and model-driven PGs, where the former is the derivative of the learned Q-value function, and the latter is that of the model-predictive return.
-
8 Oct 2020 2 repositories listed Syntology ran 0 of 13 samples · 13 unverifiedText-based games have emerged as an important test-bed for Reinforcement Learning (RL) research, requiring RL agents to combine grounded language understanding with sequential decision making.
-
15 Feb 2020 2 repositories listed Syntology ran 0 of 11 samples · 11 unverifiedWe present PDDLGym, a framework that automatically constructs OpenAI Gym environments from PDDL domains and problems.
-
20 Oct 2019 2 repositories listedIn this report, we introduce the progress to learn the policy for Malaria Control as a Reinforcement Learning problem in the KDD Cup Challenge 2019 and propose diverse solutions to deal with the limited observations…
-
5 Sep 2019 2 repositories listed Syntology ran 0 of 1 samples · 1 unverifiedThis work focuses on a specific classification problem, where the information about a sample is not readily available, but has to be acquired for a cost, and there is a per-sample budget.
-
26 Nov 2017 2 repositories listedWhile deeper convolutional networks are needed to achieve maximum accuracy in visual perception tasks, for many inputs shallower networks are sufficient.
-
2 Oct 2017 2 repositories listed Syntology ran 0 of 10 samples · 10 unverified · 10 pointer-only (licence)Our core idea is that the adversarial examples targeting at a neural network-based policy are not effective for the frame prediction model.
-
19 Jul 2017 2 repositories listedHere we introduce the "Imagination-based Planner", the first model-based, sequential decision-making agent that can learn to construct, evaluate, and execute plans.
-
11 Nov 2015 2 repositories listedWe study the problem of off-policy value evaluation in reinforcement learning (RL), where one aims to estimate the value of a new policy based on data collected by a different policy.
Syntology lines on 15 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections