Papers › Is Synthetic Dataset Reliable for Benchmarking Generalizable Person Re-Identification?

Is Synthetic Dataset Reliable for Benchmarking Generalizable Person Re-Identification?

12 Sep 2022arXiv:2209.05047archive 2025-07-28

Cuicui Kang

Recent studies show that models trained on synthetic datasets are able to achieve better generalizable person re-identification (GPReID) performance than that trained on public real-world datasets. On the other hand, due to the limitations of real-world person ReID datasets, it would also be important and interesting to use large-scale synthetic datasets as test sets to benchmark person ReID algorithms. Yet this raises a critical question: is synthetic dataset reliable for benchmarking generalizable person re-identification? In the literature there is no evidence showing this. To address this, we design a method called Pairwise Ranking Analysis (PRA) to quantitatively measure the ranking similarity and perform the statistical test of identical distributions. Specifically, we employ Kendall rank correlation coefficients to evaluate pairwise similarity values between algorithm rankings on different datasets. Then, a non-parametric two-sample Kolmogorov-Smirnov (KS) test is performed for the judgement of whether algorithm ranking correlations between synthetic and real-world datasets and those only between real-world datasets lie in identical distributions. We conduct comprehensive experiments, with ten representative algorithms, three popular real-world person ReID datasets, and three recently released large-scale synthetic datasets. Through the designed pairwise ranking analysis and comprehensive evaluations, we conclude that a recent large-scale synthetic dataset ClonedPerson can be reliably used to benchmark GPReID, statistically the same as real-world datasets. Therefore, this study guarantees the usage of synthetic datasets for both source training set and target testing set, with completely no privacy concerns from real-world surveillance data. Besides, the study in this paper might also inspire future designs of synthetic datasets.

PaperPDFCode

Code

shengcailiao/QAConv mentioned in paperpytorchMIT report

Repository list and official/mentioned flags are the archive's, frozen 2025-07-28. Reachability, where shown, is from one Syntology probe window (2026-09-16 to 2026-09-18); repositories not probed show nothing. GitHub stars are not tracked.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

BenchmarkingGeneralizable Person Re-identificationPerson Re-Identification

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Generalizable Person Re-identification ClonedPerson TransMatcher MSMT17->Rank-1 51.8 #1 of 2 Archive leaderboard report
Generalizable Person Re-identification ClonedPerson TransMatcher MSMT17->mAP 9.0 #1 of 2 Archive leaderboard report
Generalizable Person Re-identification ClonedPerson TransMatcher Market-1501->Rank-1 50.1 #1 of 2 Archive leaderboard report
Generalizable Person Re-identification ClonedPerson TransMatcher Market-1501->mAP 9.2 #1 of 2 Archive leaderboard report
Generalizable Person Re-identification ClonedPerson TransMatcher RandPerson->Rank-1 67.8 #1 of 2 Archive leaderboard report
Generalizable Person Re-identification ClonedPerson TransMatcher RandPerson->mAP 22.1 #1 of 2 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Methods

Test

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections