Browse State-of-the-Art › Task Planning
Task Planning
100 papers with code · 0 benchmarks · 1 dataset archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 100 papers with code (344 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
8 Jun 2024 5 repositories listed Syntology ran 5 of 5 samples · 0 unverified · 4 pointer-only (licence)Large Language Models (LLMs) are being deployed across various domains today.
-
12 Feb 2024 3 repositories listed Syntology ran 1 of 5 samples · 4 unverified · 4 pointer-only (licence)Visually-conditioned language models (VLMs) have seen growing adoption in applications such as visual dialogue, scene understanding, and robotic task planning; adoption that has fueled a wealth of new models such as…
-
30 Oct 2024 2 repositories listedRecent advances in LLM have been instrumental in autonomous robot control and human-robot interaction by leveraging their vast general knowledge and capabilities to understand and reason across a wide range of tasks and…
-
8 Aug 2024 2 repositories listed Syntology ran 6 of 11 samples · 5 unverifiedThe emergence of foundation models as the "brain" of EAI agents for high-level task planning has shown promising results.
-
10 Jan 2024 2 repositories listedNext, we discuss several key challenges to achieve intelligent, efficient and secure Personal LLM Agents, followed by a comprehensive survey of representative solutions to address these challenges.
-
6 Nov 2023 2 repositories listed Syntology ran 4 of 4 samples · 0 unverified · 4 pointer-only (licence)Generalizable articulated object manipulation is essential for home-assistant robots.
-
26 Aug 2023 2 repositories listedMotivated by the substantial achievements observed in Large Language Models (LLMs) in the field of natural language processing, recent research has commenced investigations into the application of LLMs for complex,…
-
1 Jul 2021 2 repositories listedAutonomous robots need to plan the tasks they carry out to fulfill their missions.
-
8 Jul 2025 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)Specifically, a user instruction is decomposed into a sequence of action proposals, each corresponding to an interaction with the GUI.
-
14 Jun 2025 1 repository listedThis survey examines the rapidly evolving field of Deep Research systems -- AI-powered applications that automate complex research workflows through the integration of large language models, advanced information…
-
8 Jun 2025 1 repository listedThe problem of relocating a set of objects to designated areas amidst movable obstacles can be framed as a Geometric Task and Motion Planning (G-TAMP) problem, a subclass of task and motion planning (TAMP).
-
3 Jun 2025 1 repository listedThe real world is messy and unstructured.
-
23 May 2025 1 repository listedTo address these challenges, we propose BEDI (Benchmark for Embodied Drone Intelligence), a systematic and standardized benchmark designed for evaluating UAV-EAs.
-
21 May 2025 1 repository listedLarge Language Model (LLM) agents can automate cybersecurity tasks and can adapt to the evolving cybersecurity landscape without re-engineering.
-
20 May 2025 1 repository listedLarge Language Models (LLMs) demonstrate strong reasoning and task planning capabilities but remain fundamentally limited in physical interaction modeling.
-
13 May 2025 1 repository listedPDDL-based symbolic task planning remains pivotal for robot autonomy yet struggles with dynamic human-robot collaboration due to scalability, re-planning demands, and delayed plan availability.
-
30 Apr 2025 1 repository listed Syntology ran 1 of 2 samples · 1 unverified · 2 pointer-only (licence)It employs three specialized agents: a routing agent, a task planning agent, and a knowledge base agent, each powered by task-specific LLMs.
-
23 Apr 2025 1 repository listedWe evaluate GoalAct on LegalAgentBench, a benchmark with multiple types of legal tasks that require the use of multiple types of tools.
-
1 Apr 2025 1 repository listed Syntology ran 0 of 5 samples · 5 unverifiedComputer use agents automate digital tasks by directly interacting with graphical user interfaces (GUIs) on computers and mobile devices, offering significant potential to enhance human productivity by completing an…
-
27 Mar 2025 1 repository listedRecent advances in language-conditioned robotic manipulation have leveraged imitation and reinforcement learning to enable robots to execute tasks from human commands.
-
21 Mar 2025 1 repository listedNew challenges inevitably arise in bimanual manipulation, necessitating not only effective task decomposition but also efficient task allocation.
-
10 Mar 2025 1 repository listedTo address this limitation, we propose a Graphormer-enhanced risk-aware task planning framework that combines LLM-based decision-making with structured safety modeling.
-
5 Mar 2025 1 repository listedThe deployment of Large Language Models (LLMs) in robotic systems presents unique safety challenges, particularly in unpredictable environments.
-
2 Mar 2025 1 repository listedTo address these limitations in dynamic environments, we propose Closed-Loop Embodied Agent (CLEA) -- a novel architecture incorporating four specialized open-source LLMs with functional decoupling for closed-loop task…
-
25 Feb 2025 1 repository listedThese relevant actions can be pre-planned to form long-horizon subtrees, significantly enhancing the planning speed and collaboration efficiency of MRBTP.
-
20 Feb 2025 1 repository listedThe model then understands this task graph as input and generates a plan for parallel execution.
-
16 Feb 2025 1 repository listedWe annotate over 2 million navigation instructions across 861 scenes and evaluate the data quality and navigation performance of trained models.
-
15 Feb 2025 1 repository listedWe evaluate D-CIPHER on multiple CTF benchmarks and LLM models via comprehensive studies to highlight the impact of our enhancements.
-
6 Feb 2025 1 repository listedFurthermore, we examine the challenges that limit adapting LLMs in MRS, including mathematical reasoning limitations, hallucination, latency issues, and the need for robust benchmarking systems.
-
6 Feb 2025 1 repository listed Syntology ran 0 of 2 samples · 2 unverifiedEffective asynchronous planning, or the ability to efficiently reason and plan over states and actions that must happen in parallel or sequentially, is essential for agents that must account for time delays, reason over…
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections