Browse State-of-the-Art › Fine-Grained Visual Categorization
Fine-Grained Visual Categorization
31 papers with code · 0 benchmarks · 5 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
5 datasets whose archive record lists this task, ordered by the archive's paper count.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Parent tasks archive 2025-07-28
Most implemented papers archive 2025-07-28
30 shown of 31 papers with code (69 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
26 Apr 2020 8 repositories listedIn this work we explore the task of instance segmentation with attribute localization, which unifies instance segmentation (detect and segment each object instance) and fine-grained visual attribute categorization…
-
20 Mar 2020 6 repositories listedTherefore, our multi-branch and multi-scale learning network(MMAL-Net) has good classification ability and robustness for images of different scales.
-
24 Apr 2020 5 repositories listedAppropriate and timely deployment of disease management depends on early disease detection.
-
8 Aug 2019 5 repositories listedViral diseases are major sources of poor yields for cassava, the 2nd largest provider of carbohydrates in Africa.
-
18 Apr 2020 2 repositories listedThis paper introduces a novel dataset FeatherV1, containing 28, 272 images of feathers categorized by 595 bird species.
-
25 Sep 2019 2 repositories listedSpecifically, we incorporate convolutional operations along edges of the tree structure, and use the routing functions in each node to determine the root-to-leaf computational paths within the tree.
-
16 Sep 2019 2 repositories listedFine-grained visual categorization is a classification task for distinguishing categories with high intra-class and small inter-class variance.
-
28 Jul 2018 2 repositories listed Syntology ran 1 of 10 samples · 9 unverifiedFine-grained visual categorization (FGVC) is challenging due in part to the fact that it is often difficult to acquire an enough number of training samples.
-
1 Oct 2017 2 repositories listedWhile the existing datasets for FGVC are mainly focused on animal breeds or man-made objects with limited labelled data, VegFru is a larger dataset consisting of vegetables and fruits which are closely associated with…
-
6 Jan 2025 1 repository listedThe discriminative concepts is utilized to guide the fine-grained representation learning.
-
23 Jul 2024 1 repository listedA major challenge in SFDA is deriving accurate categorical information for the target domain, especially when sample embeddings from different classes appear similar.
-
10 Jul 2024 1 repository listedFungiCLEF 2024 addresses the fine-grained visual categorization (FGVC) of fungi species, with a focus on identifying poisonous species.
-
10 May 2024 1 repository listed Syntology ran 12 of 13 samples · 1 unverified · 13 pointer-only (licence)To tackle this problem, we devise a Region-Aligned Proxy Learning (RAPL) framework, which comprises a Channel-wise Region Alignment (CRA) module and a Semi-Supervised Proxy Learning (SemiPL) strategy.
-
18 Apr 2024 1 repository listedOur approach utilizes an adversarial distillation framework with attention generator, mixed high-order attention distillation, and semantic feature contrast learning.
-
8 Jun 2023 1 repository listedFine-grained visual categorization (FGVC) is a challenging task due to similar visual appearances between various species.
-
4 Jun 2023 1 repository listedIn the existing FGVC datasets used in computer vision, it is generally assumed that each collected instance has fixed characteristics and the distribution of different categories is relatively balanced.
-
31 Aug 2022 1 repository listedTo address the above limitations, we propose the Structure Information Modeling Transformer (SIM-Trans) to incorporate object structure information into transformer for enhancing discriminative representation learning…
-
21 Jul 2022 1 repository listedWe thoroughly benchmark audiovisual classification performance and modality fusion experiments through the use of state-of-the-art transformer methods.
-
17 Jul 2022 1 repository listedHow- ever, the complexity of the model makes it difficult to interpret the decision-making process, and the ambiguity of the attention maps can cause incorrect correlations between image patches.
-
26 May 2022 1 repository listed Syntology ran 3 of 17 samples · 14 unverifiedInspired by this observation, we propose a network branch dedicated to magnifying the importance of small eigenvalues.
-
13 Nov 2021 1 repository listedOf those, methods based on bilinear pooling are one of the main categories for computing the interaction between deep features and have shown high effectiveness.
-
19 Aug 2021 1 repository listed Syntology ran 8 of 11 samples · 3 unverifiedUnlike most existing methods that learn visual attention based on conventional likelihood, we propose to learn the attention with counterfactual causality, which provides a tool to measure the attention quality and a…
-
6 Jul 2021 1 repository listedWe verify the effectiveness of FFVT on three benchmarks where FFVT achieves the state-of-the-art performance.
-
18 May 2021 1 repository listedThe deconstruction learning forces the model to focus on local object parts, while reconstruction learning helps in learning the correlation between the parts.
-
30 Mar 2021 1 repository listed Syntology ran 0 of 12 samples · 12 unverifiedIn order to facilitate progress in this area we present two new natural world visual classification datasets, iNat2021 and NeWT.
-
1 Jan 2021 1 repository listedThe proposed UFG image dataset and evaluation protocols is intended to serve as a benchmark platform that can advance research of visual classification from approaching human performance to beyond human ability, via…
-
28 May 2020 1 repository listedIdentifying prescription medications is a frequent task for patients and medical professionals; however, this is an error-prone task as many pills have similar appearances (e.
-
14 Jul 2019 1 repository listedFood classification is a challenging problem due to the large number of categories, high visual similarity between different foods, as well as the lack of datasets for training state-of-the-art deep models.
-
27 Nov 2018 1 repository listedWe propose FineGAN, a novel unsupervised GAN framework, which disentangles the background, object shape, and object appearance to hierarchically generate images of fine-grained object categories.
-
16 Jun 2018 1 repository listedWe propose a measure to estimate domain similarity via Earth Mover's Distance and demonstrate that transfer learning benefits from pre-training on a source domain that is similar to the target domain by this measure.
Syntology lines on 5 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections