Datasets › Stanford Cars

Stanford Cars

Introduced in 3D Object Representations for Fine-Grained Categorization6 Feb 2017 archive 2025-07-28

The Stanford Cars dataset consists of 196 classes of cars with a total of 16,185 images, taken from the rear. The data is divided into almost a 50-50 train/test split with 8,144 training images and 8,041 testing images. Categories are typically at the level of Make, Model, Year. The images are 360×240.

Source: View Independent Vehicle Make, Model and Color Recognition Using Convolutional Neural Network Image Source: https://ai.stanford.edu/~jkrause/cars/car_dataset.html

Benchmarks archive 2025-07-28

All 13 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Fine-Grained Image Classification Stanford Cars TResnet-L + PMD Accuracy 97.3% Progressive Multi-task Anti-Noise Learning and... dichao-liu/anti-noise_fgvr 83 Compare
Image Classification Stanford Cars efficient adaptive ensembling Accuracy 96.868 Efficient Adaptive Ensembling for Image Classification — 24 Compare
Prompt Engineering Stanford Cars PromptKD Harmonic mean 83.13 PromptKD: Unsupervised Prompt Distillation for... zhengli97/promptkd 14 Compare
Continual Learning Stanford Cars (Fine-grained 6 Tasks) CPG Accuracy 92.80 Compacting, Picking and Growing for Unforgetting... ivclab/CPG +1 6 Compare
Few-Shot Image Classification Stanford Cars 5-way (1-shot) MATANet Accuracy 73.15 Multi-scale Adaptive Task Attention Network for Few-Shot Learning — 6 Compare
Few-Shot Image Classification Stanford Cars 5-way (5-shot) MATANet Accuracy 91.89 Multi-scale Adaptive Task Attention Network for Few-Shot Learning — 6 Compare
Image Clustering Stanford Cars TURTLE (CLIP + DINOv2) Accuracy 0.646 Let Go of Your Labels with Unsupervised Transfer mlbio-epfl/turtle 5 Compare
Image Generation Stanford Cars Projected GANs FID 2.09 Projected GANs Converge Faster autonomousvision/projected_gan +2 4 Compare
Neural Architecture Search Stanford Cars NAT-M4 Accuracy (%) 92.9 Neural Architecture Transfer human-analysis/neural-architecture-transfer +1 4 Compare
Few-Shot Learning Stanford Cars SaSPA + CAL 4-shot Accuracy 66.7 Advancing Fine-Grained Classification by Structure and... eyalmichaeli/saspa-aug 3 Compare
Learning with coarse labels Stanford Cars MaskCon Recall@1 45.53 MaskCon: Masked Contrastive Learning for Coarse-Labelled Dataset MrChenFeng/MaskCon_CVPR2023 2 Compare
Zero-Shot Learning Stanford Cars ZLaP* Accuracy 71.8 Label Propagation for Zero-shot Classification with... vladan-stojnic/zlap 2 Compare
Transductive Zero-Shot Classification Stanford Cars ZLaP Accuracy 72.1 Label Propagation for Zero-shot Classification with... vladan-stojnic/zlap 1 Compare

Papers archive 2025-07-28

30 shown of 123 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 790. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics 1 1 3 Jul 2025 not harvested
MMRL++: Parameter-Efficient and Interaction-Aware Representation Learning for Vision-Language Models 1 1 15 May 2025 not harvested
MMRL: Multi-Modal Representation Learning for Vision-Language Models 1 1 11 Mar 2025 ran 0 of 1 samples (1 unverified)
Interweaving Insights: High-Order Feature Interaction for Fine-Grained Visual Recognition 1 1 20 Oct 2024 not harvested
Stochastic Subsampling With Average Pooling 0 1 25 Sep 2024 not harvested
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling 2 1 27 Aug 2024 not harvested
Multi-Granularity Part Sampling Attention for Fine-Grained Visual Classification 1 1 16 Aug 2024 not harvested
Advancing Fine-Grained Classification by Structure and Subject Preserving Augmentation 1 2 20 Jun 2024 ran 0 of 1 samples (1 unverified)
Let Go of Your Labels with Unsupervised Transfer 1 1 11 Jun 2024 ran 3 of 4 samples (1 unverified; 4 pointer-only for licence)
Label Propagation for Zero-shot Classification with Vision-Language Models 1 3 5 Apr 2024 ran 1 of 2 samples (1 unverified)
Prompt Learning via Meta-Regularization 1 1 1 Apr 2024 ran 4 of 7 samples (3 unverified)
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs 3 4 28 Mar 2024 ran 3 of 3 samples (0 unverified)
Context-Semantic Quality Awareness Network for Fine-Grained Visual Categorization 0 1 15 Mar 2024 not harvested
PromptKD: Unsupervised Prompt Distillation for Vision-Language Models 1 1 5 Mar 2024 not harvested
Progressive Multi-task Anti-Noise Learning and Distilling Frameworks for Fine-grained Vehicle Recognition 1 1 25 Jan 2024 not harvested
Learning Hierarchical Prompt with Structured Linguistic Knowledge for Vision-Language Models 2 1 11 Dec 2023 ran 3 of 7 samples (4 unverified)
DePT: Decoupled Prompt Tuning 1 1 14 Sep 2023 ran 4 of 8 samples (4 unverified; 8 pointer-only for licence)
Read-only Prompt Optimization for Vision-Language Few-shot Learning 1 1 29 Aug 2023 ran 4 of 6 samples (2 unverified)
PCNN: Probable-Class Nearest-Neighbor Explanations Improve Fine-Grained Image Classification Accuracy for AIs and Humans 2 1 25 Aug 2023 not harvested
Multiscale patch-based feature graphs for image classification 1 1 8 Aug 2023 not harvested
Self-regulating Prompts: Foundational Model Adaptation without Forgetting 2 1 13 Jul 2023 ran 7 of 21 samples (14 unverified)
Consistency-guided Prompt Learning for Vision-Language Models 2 1 1 Jun 2023 ran 3 of 7 samples (4 unverified)
Is Synthetic Data From Diffusion Models Ready for Knowledge Distillation? 1 1 22 May 2023 ran 1 of 1 samples (0 unverified; 1 pointer-only for licence)
Understanding Gaussian Attention Bias of Vision Transformers Using Effective Receptive Fields 1 2 8 May 2023 not harvested
MaskCon: Masked Contrastive Learning for Coarse-Labelled Dataset 1 1 22 Mar 2023 ran 0 of 2 samples (2 unverified)
Learn from Each Other to Classify Better: Cross-layer Mutual Attention Learning for Fine-grained Visual Classification 2 1 22 Mar 2023 not harvested
Part-guided Relational Transformers for Fine-grained Visual Recognition 1 1 28 Dec 2022 not harvested
Learning Domain Invariant Prompt for Vision-Language Models 1 1 8 Dec 2022 not harvested
Penalizing the Hard Example But Not Too Much: A Strong Baseline for Fine-Grained Visual Classification 1 1 21 Nov 2022 not harvested
Helpful or Harmful: Inter-Task Association in Continual Learning 1 1 23 Oct 2022 not harvested

The full list of 123 is in the JSON twin.

Dataset loaders archive 2025-07-28

3 loaders as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Custom (non-commercial)

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • Stanford Cars
  • Stanford Cars 5-way (1-shot)
  • Stanford Cars 5-way (5-shot)
  • Stanford Cars (Fine-grained 6 Tasks)

4 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections