Browse State-of-the-Art › Object Discovery
Object Discovery
96 papers with code · 0 benchmarks · 3 datasets archive 2025-07-28
Object Discovery is the task of identifying previously unseen objects.
Source: Unsupervised Object Discovery and Segmentation of RGBD-images
Description from the archive archive 2025-07-28.
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
3 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 96 papers with code (210 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
26 Jun 2020 8 repositories listed Syntology ran 14 of 21 samples · 7 unverified · 7 pointer-only (licence)Learning object-centric representations of complex scenes is a promising step towards enabling efficient abstract reasoning from low-level perceptual features.
-
28 Sep 2023 6 repositories listed Syntology ran 4 of 20 samples · 16 unverified · 2 pointer-only (licence)Transformers have recently emerged as a powerful tool for learning visual representations.
-
15 Aug 2021 6 repositories listed Syntology ran 0 of 4 samples · 4 unverifiedIn this paper, we identify that the problem is that the binary classifiers in existing proposal methods tend to overfit to the training categories.
-
22 Jan 2019 5 repositories listed Syntology ran 1 of 7 samples · 6 unverifiedThe ability to decompose scenes in terms of abstract building blocks is crucial for general intelligence.
-
6 Apr 2018 4 repositories listedWe propose an end-to-end-trainable attention module for convolutional neural network (CNN) architectures built for image classification.
-
23 Nov 2016 4 repositories listed Syntology ran 0 of 3 samples · 3 unverifiedOur key contribution is the collection of a large-scale dataset consisting of 150K human-played games with a total of 800K visual question-answer pairs on 66K images.
-
24 Sep 2023 3 repositories listedUsing our dataset and annotations, we release benchmarks for 3D object detection and 3D semantic segmentation using established metrics.
-
12 Mar 2024 2 repositories listed Syntology ran 6 of 9 samples · 3 unverifiedIn this paper, we introduce VoteCut, an innovative method for unsupervised object discovery that leverages feature representations from multiple self-supervised models.
-
27 Mar 2023 2 repositories listed Syntology ran 18 of 29 samples · 11 unverified · 29 pointer-only (licence)Object discovery -- separating objects from the background without manual labels -- is a fundamental open challenge in computer vision.
-
9 Feb 2023 2 repositories listedAutomatically discovering composable abstractions from raw perceptual data is a long-standing challenge in machine learning.
-
20 Mar 2022 2 repositories listed Syntology ran 4 of 11 samples · 7 unverifiedPrevious advances in object tracking mostly reported on favorable illumination circumstances while neglecting performance at nighttime, which significantly impeded the development of related aerial robot applications.
-
9 Mar 2022 2 repositories listedTo this end, we (i) exhaustively evaluate common meta-learning techniques on these tasks, and (ii) quantitatively analyze the effect of various deep learning techniques commonly used in recent meta-learning algorithms…
-
7 Oct 2021 2 repositories listedThe ability to decompose scenes into their object components is a desired property for autonomous agents, allowing them to reason and act in their surroundings.
-
29 Sep 2021 2 repositories listedWe also show that training a class-agnostic detector on the discovered objects boosts results by another 7 points.
-
30 Jul 2019 2 repositories listedGenerative latent-variable models are emerging as promising tools in robotics and reinforcement learning.
-
22 May 2019 2 repositories listed Syntology ran 0 of 6 samples · 6 unverifiedData efficiency and robustness to task-irrelevant perturbations are long-standing challenges for deep reinforcement learning algorithms.
-
2 Oct 2018 2 repositories listedThis paper is concerned with the training of recurrent neural networks as goal-oriented dialog agents using reinforcement learning.
-
2 Jul 2025 1 repository listedPruning is widely used to reduce the complexity of deep learning models, but its effects on interpretability and representation learning remain poorly understood.
-
6 May 2025 1 repository listedFor threshold units, this leads to Hopfield associative memory model, and for oscillatory units it yields a specific form of generalized Kuramoto model.
-
9 Apr 2025 1 repository listed Syntology ran 0 of 8 samples · 8 unverifiedHowever, with recent sample-efficient segmentation models, we can separate objects in the pixel space and encode them independently.
-
27 Feb 2025 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedDecomposing visual scenes into objects, as humans do, facilitates modeling object relations and dynamics.
-
8 Jan 2025 1 repository listedThis work aims at leveraging 2D visual cues for efficient autonomous exploration, addressing the limitations of extracting goal poses from a 3D map.
-
17 Nov 2024 1 repository listedReconstructing compositional 3D representations of scenes, where each object is represented with its own 3D model, is a highly desirable capability in robotics and augmented reality.
-
4 Nov 2024 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedVisualizations show that our GDR captures better object separability.
-
17 Oct 2024 1 repository listed Syntology ran 3 of 3 samples · 0 unverified · 3 pointer-only (licence)It has long been known in both neuroscience and AI that ``binding'' between neurons leads to a form of competitive learning where representations are compressed in order to represent more abstract concepts in deeper…
-
24 Jul 2024 1 repository listedLocalizing objects in an unsupervised manner poses significant challenges due to the absence of key visual information such as the appearance, type and number of objects, as well as the lack of labeled object classes…
-
3 Jul 2024 1 repository listedConsidering the adopted bidirectional alignment will also weaken the anchor image activation if appropriate constraints are missing, we propose a self-supervised regularization module to maintain the reliable activation…
-
21 Jun 2024 1 repository listed Syntology ran 5 of 7 samples · 2 unverifiedWe demonstrate the effectiveness of DiPEx through extensive class-agnostic OD and OOD-OD experiments on MS-COCO and LVIS, surpassing other prompting methods by up to 20.
-
13 Jun 2024 1 repository listedMoreover, our analysis substantiates that our method exhibits the capability to dynamically adapt the slot number according to each instance's complexity, offering the potential for further exploration in slot attention…
-
2 Jun 2024 1 repository listed Syntology ran 7 of 11 samples · 4 unverified3D-NOD is further extended with an Enrichment strategy that significantly enriches the novel object distribution in the training scenes, and then enhances the model's ability to localize more novel objects.
Syntology lines on 15 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections