Browse State-of-the-Art › Novel Object Detection
Novel Object Detection
24 papers with code · 1 benchmark · 1 dataset archive 2025-07-28
Novel Object Detection is a challenging task introduced by Fomenko et.al. in their paper "Learning to Discover and Detect Objects". The goal in this task is to measure mAP performance on known as well as novel classes, where the known classes correspond to the 80 COCO classes, and the novel classes are the remaining 1123 classes from LVIS dataset. Thus, during training the model can only be trained with annotations from COCO dataset, but during evaluation/inference it is expected to BOTH classify and detect objects belonging to ALL the classes in the LVIS dataset.
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
1 leaderboard table shown for this task, 1 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted.
| Dataset | Best model (first row in archive order) | Paper | Code | Syntology | Compare |
|---|---|---|---|---|---|
| LVIS v1.0 val (5 rows) | Cooperative Foundational Models | Enhancing Novel Object Detection via Cooperative Foundational Models | code | — | Compare |
Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
1 dataset whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
24 shown of 24 papers with code (53 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
9 Jun 2024 2 repositories listedOur contributions are as follows: 1) We propose that the ODMamba backbone introduce a \textbf{S}tate \textbf{S}pace \textbf{M}odel (\textbf{SSM}) with linear complexity to address the quadratic complexity of…
-
12 Jun 2020 2 repositories listedDeploying deep learning models on embedded systems has been challenging due to limited computing resources.
-
29 Nov 2018 2 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedThis paper proposes a novel object detection framework named Grid R-CNN, which adopts a grid guided localization mechanism for accurate object detection.
-
5 Jul 2024 1 repository listedMeanwhile, the Advanced Assisted Fusion (AAF) module deeply embedded within the neck conveys a more diverse range of gradient information to the output layer.
-
14 Mar 2024 1 repository listedIn this paper we present YOLOX-ViT, a novel object detection model, and investigate the efficacy of knowledge distillation for model size reduction without sacrificing performance.
-
15 Jan 2024 1 repository listedHowever, the class-level prototypes are difficult to precisely generate, and they also lack detailed information, leading to instability in performance.
-
19 Nov 2023 1 repository listedWe present a novel approach to transform existing closed-set detectors into open-set detectors.
-
3 Oct 2023 1 repository listedMFAD excels in both simple and complex anomaly detection scenarios.
-
2 Oct 2023 1 repository listedWe refer to this approach as the self-training strategy, which enhances recall and accuracy for novel classes without requiring extra annotations, datasets, and re-training.
-
17 Sep 2023 1 repository listedThe ability to detect objects in all lighting (i.
-
24 Jun 2023 1 repository listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)In fact, our experiments show that GLIP, the state-of-the-art vision-language model for object detection, often disregards contextual information in the language descriptions and instead relies heavily on detecting…
-
19 Oct 2022 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedWe then train our network to learn to classify each RoI, either as one of the known classes, seen in the source dataset, or one of the novel classes, with a long-tail distribution constraint on the class assignments,…
-
11 Jul 2022 1 repository listedA critical object detection task is finetuning an existing model to detect novel objects, but the standard workflow requires bounding box annotations which are time-consuming and expensive to collect.
-
24 Oct 2021 1 repository listedDue to the success of Bidirectional Encoder Representations from Transformers (BERT) in natural language process (NLP), the multi-head attention transformer has been more and more prevalent in computer-vision researches…
-
19 Aug 2021 1 repository listed Syntology ran 4 of 4 samples · 0 unverifiedIn this paper, we study the problem of Novel Class Discovery (NCD).
-
26 Jul 2021 1 repository listedAccess to high resolution satellite imagery has dramatically increased in recent years as several new constellations have entered service.
-
15 May 2021 1 repository listedThe model achieves a (COCO-style) average precision of 0.
-
1 Mar 2021 1 repository listedThus, we develop a new framework of few-shot object detection with universal prototypes ({FSOD}^{up}) that owns the merit of feature generalization towards novel objects.
-
6 Feb 2021 1 repository listedHere, we introduce a novel open-world semi-supervised learning setting that formalizes the notion that novel classes may appear in the unlabeled test data.
-
4 Mar 2020 1 repository listedWe have taken an incremental approach in reaching our final proposed method through detailed evaluation and comparison with baselines using our constructed SVSO (Street View Signboard Objects) signboard dataset…
-
27 May 2019 1 repository listedIn this work, we propose to exploit the natural correlation in narrations and the visual presence of objects in video, to learn an object detector and retrieval without any manual labeling involved.
-
3 Mar 2019 1 repository listedThis paper presents a novel object detection network (CAD-Net) that exploits attention-modulated features as well as global and local contexts to address the new challenges in detecting objects from remote sensing…
-
14 Mar 2017 1 repository listedWe present a novel object detection pipeline for localization and recognition in three dimensional environments.
-
9 Apr 2016 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedTemporal and contextual information of videos are not fully investigated and utilized.
Syntology lines on 5 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections