Datasets › Food-101

Food-101

Introduced in Food-101 - Mining Discriminative Components with Random Forests1 Jan 2014 archive 2025-07-28

The Food-101 dataset consists of 101 food categories with 750 training and 250 test images per category, making a total of 101k images. The labels for the test images have been manually cleaned, while the training set contains some noise.

Source: Combining Weakly and Webly Supervised Learning for Classifying Food Images Image Source: https://data.vision.ee.ethz.ch/cvl/datasets_extra/food-101/

Benchmarks archive 2025-07-28

All 14 leaderboards whose dataset resolves to this page shown (sort by any header). "First row" is the archive's own first row at snapshot, in the archive's row order; nothing here re-ranks and metric direction is not asserted.

First row (archive order)PaperCode
Fine-Grained Image Classification Food-101 CAP Accuracy 98.6 Context-aware Attentional Pooling (CAP) for Fine-grained... ArdhenduBehera/cap 15 Compare
Prompt Engineering Food-101 PromptKD Harmonic mean 93.05 PromptKD: Unsupervised Prompt Distillation for... zhengli97/promptkd 13 Compare
Image Classification Food-101 Bamboo (ViTB/16) Accuracy (%) 92.9 Bamboo: Building Mega-Scale Vision Dataset Continually... zhangyuanhan-ai/bamboo +1 11 Compare
Neural Architecture Search Food-101 NAT-M4 Accuracy (%) 89.4 Neural Architecture Transfer human-analysis/neural-architecture-transfer +1 5 Compare
Zero-Shot Transfer Image Classification Food-101 MAWS (ViT-2B) Top 1 Accuracy 96.2 The effectiveness of MAE pre-pretraining for... facebookresearch/maws 5 Compare
Learning with noisy labels Food-101 LongReMix Accuracy (% ) 86.42 LongReMix: Robust Learning with High Confidence Samples... filipe-research/LongReMix 3 Compare
Multimodal Text and Image Classification Food-101 Early Fusion (Bert + InceptionV3) Accuracy (%) 92.5 Image and Text fusion for UPMC Food-101 \\using BERT and CNNs artelab/Image-and-Text-fusion-for-UPMC-Food-101-using-BERT-and-CNNs 2 Compare
Zero-Shot Learning Food-101 ZLaP* Accuracy 87.9 Label Propagation for Zero-shot Classification with... vladan-stojnic/zlap 2 Compare
Document Text Classification Food-101 Bert Accuracy (%) 84.41 Image and Text fusion for UPMC Food-101 \\using BERT and CNNs artelab/Image-and-Text-fusion-for-UPMC-Food-101-using-BERT-and-CNNs 1 Compare
Few-Shot Learning food101 Variational Prompt Tuning Harmonic mean 91.57 Bayesian Prompt Learning for Image-Language Model Generalization saic-fi/bayesian-prompt-learning 1 Compare
Image Clustering Food-101 TURTLE (CLIP + DINOv2) Accuracy 92.2 Let Go of Your Labels with Unsupervised Transfer mlbio-epfl/turtle 1 Compare
Image Compression Food-101 Lossyless Compressor Bit rate 1270 Lossy Compression for Lossless Prediction YannDubs/lossyless 1 Compare
Transductive Zero-Shot Classification Food-101 ZLaP Accuracy 87.9 Label Propagation for Zero-shot Classification with... vladan-stojnic/zlap 1 Compare
Classification food101 no rows — — 0 Compare

Papers archive 2025-07-28

30 shown of 45 papers with a leaderboard row on this dataset's benchmarks, newest first. The archive's own "papers using this dataset" list was never published, so this is the benchmark-backed subset; the archive's count for this dataset is 805. The Syntology column is from Syntology's graph (read 2026-09-24), stated per sample; it is not part of any archive number.

DateSamples run Syntology
Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics 1 1 3 Jul 2025 not harvested
SST: Self-training with Self-adaptive Thresholding for Semi-supervised Learning 0 4 31 May 2025 not harvested
MMRL++: Parameter-Efficient and Interaction-Aware Representation Learning for Vision-Language Models 1 1 15 May 2025 not harvested
MMRL: Multi-Modal Representation Learning for Vision-Language Models 1 1 11 Mar 2025 ran 0 of 1 samples (1 unverified)
PSSCL: A progressive sample selection framework with contrastive loss designed for noisy labels 1 1 18 Dec 2024 not harvested
HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling 2 1 27 Aug 2024 not harvested
Let Go of Your Labels with Unsupervised Transfer 1 1 11 Jun 2024 ran 3 of 4 samples (1 unverified; 4 pointer-only for licence)
Label Propagation for Zero-shot Classification with Vision-Language Models 1 3 5 Apr 2024 ran 1 of 2 samples (1 unverified)
Prompt Learning via Meta-Regularization 1 1 1 Apr 2024 ran 4 of 7 samples (3 unverified)
PromptKD: Unsupervised Prompt Distillation for Vision-Language Models 1 1 5 Mar 2024 not harvested
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters 2 1 6 Feb 2024 ran 1 of 6 samples (5 unverified)
InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks 2 1 21 Dec 2023 ran 2 of 2 samples (0 unverified; 2 pointer-only for licence)
Learning Hierarchical Prompt with Structured Linguistic Knowledge for Vision-Language Models 2 1 11 Dec 2023 ran 3 of 7 samples (4 unverified)
Dining on Details: LLM-Guided Expert Networks for Fine-Grained Food Recognition 0 1 29 Oct 2023 not harvested
DePT: Decoupled Prompt Tuning 1 1 14 Sep 2023 ran 4 of 8 samples (4 unverified; 8 pointer-only for licence)
Read-only Prompt Optimization for Vision-Language Few-shot Learning 1 1 29 Aug 2023 ran 4 of 6 samples (2 unverified)
Self-regulating Prompts: Foundational Model Adaptation without Forgetting 2 1 13 Jul 2023 ran 7 of 21 samples (14 unverified)
Balanced Mixture of SuperNets for Learning the CNN Pooling Architecture 1 1 21 Jun 2023 not harvested
Consistency-guided Prompt Learning for Vision-Language Models 2 1 1 Jun 2023 ran 3 of 7 samples (4 unverified)
Your Diffusion Model is Secretly a Zero-Shot Classifier 4 1 28 Mar 2023 ran 2 of 2 samples (0 unverified; 1 pointer-only for licence)
EVA-CLIP: Improved Training Techniques for CLIP at Scale 4 1 27 Mar 2023 ran 0 of 4 samples (4 unverified)
The effectiveness of MAE pre-pretraining for billion-scale pretraining 1 1 23 Mar 2023 not harvested
Fine-Grained Classification with Noisy Labels 0 1 4 Mar 2023 not harvested
Learning Domain Invariant Prompt for Vision-Language Models 1 1 8 Dec 2022 not harvested
Learning Multi-Subset of Classes for Fine-Grained Food Recognition 1 2 10 Oct 2022 not harvested
MaPLe: Multi-modal Prompt Learning 3 1 6 Oct 2022 ran 4 of 6 samples (2 unverified)
Bayesian Prompt Learning for Image-Language Model Generalization 1 1 5 Oct 2022 ran 4 of 6 samples (2 unverified)
A Continual Development Methodology for Large-scale Multitask Dynamic ML Systems 1 1 15 Sep 2022 not harvested
TransBoost: Improving the Best ImageNet Performance using Deep Transduction 1 1 26 May 2022 not harvested
Bamboo: Building Mega-Scale Vision Dataset Continually with Human-Machine Synergy 2 1 15 Mar 2022 ran 3 of 3 samples (0 unverified; 3 pointer-only for licence)

The full list of 45 is in the JSON twin.

Dataset loaders archive 2025-07-28

12 loaders as listed in the archive; links are outbound and not re-checked here.

Tasks archive 2025-07-28

License archive 2025-07-28

Unknown

Modalities archive 2025-07-28

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • Food-101
  • food101

2 variant names, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections