Browse State-of-the-Art › Text based Person Search
Text based Person Search
18 papers with code · 0 benchmarks · 2 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
2 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
18 shown of 18 papers with code (37 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
16 Jul 2022 2 repositories listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)In PGU, we adopt a set of shared and learnable prototypes as the queries to extract diverse and semantically aligned features for both modalities in the granularity-unified feature space, which further promotes the ReID…
-
8 Jan 2021 2 repositories listedSecondly, a BERT with locality-constrained attention is proposed to obtain representations of descriptions at different scales.
-
30 Dec 2024 1 repository listedPrior works adopt image and text encoders pre-trained on unimodal data to extract global and local features from image and text respectively, and then global-local alignment is achieved explicitly.
-
26 Nov 2024 1 repository listedTo enable the training and evaluation of this new task, we construct a large-scale image-text Pedestrian Anomaly Behavior (PAB) benchmark, featuring a broad spectrum of actions, e.
-
5 Jul 2024 1 repository listedThe task is that of retrieving one or more images of a specific individual based on a textual description.
-
16 Apr 2024 1 repository listedIn text-based person search endeavors, data generation has emerged as a prevailing practice, addressing concerns over privacy preservation and the arduous task of manual annotation.
-
15 Nov 2023 1 repository listed Syntology ran 7 of 7 samples · 0 unverified · 7 pointer-only (licence)Moreover, we propose a proximity data generation (PDG) module to automatically produce more diverse data for cross-modal training.
-
19 Aug 2023 1 repository listedTPBS, as a fine-grained cross-modal retrieval task, is also facing the rise of research on the CLIP-based TBPS.
-
23 May 2023 1 repository listed Syntology ran 3 of 7 samples · 4 unverifiedRA offsets the overfitting risk by introducing a novel positive relation detection task (i.
-
12 Apr 2023 1 repository listedRecent researches on unsupervised person re-identification~(reID) have demonstrated that pre-training on unlabeled person images achieves superior performance on downstream reID tasks than pre-training on ImageNet.
-
26 Nov 2022 1 repository listedTo implement this task, one needs to extract multi-scale features from both image and text domains, and then perform the cross-modal alignment.
-
16 Nov 2022 1 repository listedSpecifically, we improve the interpretability of text features by providing them with consistent semantic information with image features to achieve the alignment of text and describe image region features.
-
4 Nov 2022 1 repository listedText-based person search aims to associate pedestrian images with natural language descriptions.
-
19 Oct 2022 1 repository listed Syntology ran 10 of 12 samples · 2 unverifiedSecondly, cross-grained feature refinement (CFR) and fine-grained correspondence discovery (FCD) modules are proposed to establish the cross-grained and fine-grained interactions between modalities, which can filter out…
-
13 Dec 2021 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedIn this paper, we propose a semantic-aligned embedding method for text-based person search, in which the feature alignment across modalities is achieved by automatically learning the semantic-aligned visual features and…
-
20 Oct 2021 1 repository listedFirstly, to fully utilize the existing small-scale benchmarking datasets for more discriminative feature learning, we introduce a cross-modal momentum contrastive learning framework to enrich the training data for a…
-
27 Sep 2021 1 repository listedFinding target persons in full scene images with a query of text description has important practical applications in intelligent video surveillance.
-
25 May 2021 1 repository listedText-based person search is a sub-task in the field of image retrieval, which aims to retrieve target person images according to a given textual description.
Syntology lines on 5 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections