Browse State-of-the-Art › NavSim
NavSim
14 papers with code · 1 benchmark · 0 datasets archive 2025-07-28
Data-Driven Non-Reactive Autonomous Vehicle Benchmark.
Evaluates autonomous driving stacks that produce waypoints with a static dataset using the PDM-Score metric. The PDM-score metric performs a pseudo-simulation by rolling out the trajectory and simulating all other actors via log-replay. This results in an open-loop evaluation that correlates with closed-loop performance.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| OpenScene (29 rows) | DriveSuprim | — | — | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
14 shown of 14 papers with code (26 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
11 Jun 2024 3 repositories listedWe propose Hydra-MDP, a novel paradigm employing multiple teachers in a teacher-student model.
-
31 May 2022 3 repositories listed Syntology ran 1 of 5 samples · 4 unverifiedAt the time of submission, TransFuser outperforms all prior work on the CARLA leaderboard in terms of driving score by a large margin.
-
21 Jun 2024 2 repositories listed Syntology ran 1 of 1 samples · 0 unverifiedOn a large set of challenging scenarios, we observe that simple methods with moderate compute requirements such as TransFuser can match recent large-scale end-to-end driving architectures such as UniAD.
-
30 Jun 2025 1 repository listedDiffusion models have demonstrated exceptional visual quality in video generation, making them promising for autonomous driving world modeling.
-
7 Jun 2025 1 repository listedGTRS consists of three complementary innovations: (1) a diffusion-based trajectory generator that produces diverse fine-grained proposals; (2) a vocabulary generalization technique that trains a scorer on super-dense…
-
4 Jun 2025 1 repository listedOur method then assigns a higher importance to synthetic observations that best match the AV's likely behavior using a novel proximity-based weighting scheme.
-
21 May 2025 1 repository listedEnd-to-end (E2E) autonomous driving systems offer a promising alternative to traditional modular pipelines by reducing information loss and error accumulation, with significant potential to enhance both mobility and…
-
2 Apr 2025 1 repository listed Syntology ran 7 of 11 samples · 4 unverifiedTherefore, we propose an end-to-end driving framework WoTE, which leverages a BEV World model to predict future BEV states for Trajectory Evaluation.
-
15 Mar 2025 1 repository listedHydra-NeXt surpasses the previous state-of-the-art by 22.
-
7 Mar 2025 1 repository listed Syntology ran 0 of 7 samples · 7 unverifiedFurthermore, GoalFlow employs an efficient generative method, Flow Matching, to generate multimodal trajectories, and incorporates a refined scoring mechanism to select the optimal trajectory from the candidates.
-
22 Nov 2024 1 repository listed Syntology ran 1 of 8 samples · 7 unverifiedHowever, the numerous denoising steps in the robotic diffusion policy and the more dynamic, open-world nature of traffic scenes pose substantial challenges for generating diverse driving actions at a real-time speed.
-
12 Jun 2024 1 repository listed Syntology ran 1 of 4 samples · 3 unverifiedSpecifically, our framework \textbf{LAW} uses a LAtent World model to predict future latent features based on the predicted ego actions and the latent feature of the current frame.
-
20 Feb 2024 1 repository listed Syntology ran 2 of 3 samples · 1 unverified · 2 pointer-only (licence)Learning a human-like driving policy from large-scale driving demonstrations is promising, but the uncertainty and non-deterministic nature of planning make it challenging.
-
20 Dec 2022 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedOriented at this, we revisit the key components within perception and prediction, and prioritize the tasks such that all these tasks contribute to planning.
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections