Browse State-of-the-Art › Image Classification

Image Classification

4,702 papers with code · 166 benchmarks · 283 datasets archive 2025-07-28

AdversarialComputer Vision

Image Classification is a fundamental task in vision recognition that aims to understand and categorize an image as a whole under a specific label. Unlike object detection, which involves classification and location of multiple objects within an image, image classification typically pertains to single-object images. When the classification becomes highly detailed or reaches instance-level, it is often referred to as image retrieval, which also involves finding similar images in a large database.

Source: Metamorphic Testing for Object Detection Systems

Description from the archive archive 2025-07-28; Papers-with-Code links inside it are rewritten to this site.

Benchmarks archive 2025-07-28

177 leaderboard tables shown for this task, 166 with rows (a “benchmark” on this site is a table with at least one row, as on /sota), ordered by row count. “Best model” is the first row in the archive's own order at snapshot; nothing is re-ranked here and metric direction is not recorded in the archive. PwC's Trend sparklines are not in the archive, so that column is omitted. 10 shown of 177 until expanded.

DatasetBest model (first row in archive order)PaperCodeSyntologyCompare
ImageNet (1,060 rows) CoCa (finetuned) CoCa: Contrastive Captioners are Image-Text Foundation Models code Syntology ran 9 of 17 samples · 8 unverified Compare
CIFAR-10 (265 rows) ViT-H/14 An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale code Syntology ran 281 of 419 samples · 138 unverified Compare
CIFAR-100 (211 rows) EffNet-L2 (SAM) Sharpness-Aware Minimization for Efficiently Improving Generalization code Syntology ran 8 of 20 samples · 12 unverified Compare
STL-10 (117 rows) µ2Net+ (ViT-L/16) A Continual Development Methodology for Large-scale Multitask... code — Compare
ObjectNet (106 rows) CoCa CoCa: Contrastive Captioners are Image-Text Foundation Models code Syntology ran 9 of 17 samples · 8 unverified Compare
MNIST (81 rows) Branching/Merging CNN + Homogeneous Vector Capsules No Routing Needed Between Capsules code — Compare
SVHN (62 rows) E2E-M3 Rethinking Recurrent Neural Networks and Other Improvements for... code — Compare
iNaturalist 2018 (60 rows) OmniVec2 OmniVec2 - A Novel Transformer based Network for Large Scale... — — Compare
ImageNet ReaL (57 rows) Baseline (ViT-G/14) Model soups: averaging weights of multiple fine-tuned models... code Syntology ran 5 of 17 samples · 12 unverified Compare
Flowers-102 (52 rows) CCT-14/7x2 Escaping the Big Data Paradigm with Compact Transformers code Syntology ran 3 of 6 samples · 3 unverified Compare
Clothing1M (51 rows) LRA-diffusion (CC) Label-Retrieval-Augmented Diffusion Models for Learning from Noisy Labels code Syntology ran 7 of 9 samples · 2 unverified Compare
mini WebVision 1.0 (47 rows) LRA-diffusion (CLIP ViT) Label-Retrieval-Augmented Diffusion Models for Learning from Noisy Labels code Syntology ran 7 of 9 samples · 2 unverified Compare
Fashion-MNIST (34 rows) PreAct-ResNet18 + FMix FMix: Enhancing Mixed Sample Data Augmentation code Syntology ran 8 of 8 samples · 0 unverified Compare
VTAB-1k (34 rows) ALIGN (50 hypers/task) Scaling Up Visual and Vision-Language Representation Learning With... code Syntology ran 8 of 10 samples · 2 unverified Compare
ImageNet V2 (33 rows) Model soups (BASIC-L) Model soups: averaging weights of multiple fine-tuned models... code Syntology ran 5 of 17 samples · 12 unverified Compare
Kuzushiji-MNIST (26 rows) KMNIST-Tiny Efficient Global Neural Architecture Search code — Compare
Stanford Cars (24 rows) efficient adaptive ensembling Efficient Adaptive Ensembling for Image Classification — — Compare
Tiny ImageNet Classification (23 rows) Astroformer Astroformer: More Data Might not be all you need for Classification code Syntology ran 1 of 1 samples · 0 unverified Compare
iNaturalist 2019 (22 rows) Hiera-H (448px) Hiera: A Hierarchical Vision Transformer without the Bells-and-Whistles code Syntology ran 0 of 6 samples · 6 unverified Compare
OmniBenchmark (22 rows) NOAH-ViTB/16 Neural Prompt Search code Syntology ran 6 of 9 samples · 3 unverified Compare
EMNIST-Balanced (20 rows) EMNIST-mobile Efficient Global Neural Architecture Search code — Compare
RESISC45 (20 rows) ResNet50 In-domain representation learning for remote sensing code — Compare
DF20 (19 rows) ViT-Large/16 (384) Danish Fungi 2020 -- Not Just Another Image Recognition Dataset code — Compare
DF20 - Mini (19 rows) ViT-Large/16 (384) Danish Fungi 2020 -- Not Just Another Image Recognition Dataset code — Compare
iNaturalist (19 rows) AIMv2-3B (448 res) Multimodal Autoregressive Pre-training of Large Vision Encoders code Syntology ran 0 of 3 samples · 3 unverified Compare
ColonINST-v1 (Seen) (17 rows) ColonGPT (w/ LoRA, w/o extra data) Frontiers in Intelligent Colonoscopy code Syntology ran 0 of 6 samples · 6 unverified Compare
ColonINST-v1 (Unseen) (17 rows) ColonGPT (w/ LoRA, w/o extra data) Frontiers in Intelligent Colonoscopy code Syntology ran 0 of 6 samples · 6 unverified Compare
WebVision-1000 (16 rows) MAM (ViT-B/16) Improving Image Recognition by Retrieving from Web-Scale Image-Text Data — — Compare
EuroSAT (15 rows) DeepEnsembling Deep Ensembling of Multiband Images for Earth Remote Sensing and... code — Compare
Places205 (15 rows) InternImage-H InternImage: Exploring Large-Scale Vision Foundation Models with... code Syntology ran 2 of 4 samples · 2 unverified Compare
DTD (11 rows) Linear FT(ViT-L/14) Task Arithmetic in the Tangent Space: Improved Editing of... code Syntology ran 1 of 4 samples · 3 unverified Compare
EMNIST-Letters (11 rows) WaveMixLite-112/16 WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
Food-101 (11 rows) Bamboo (ViTB/16) Bamboo: Building Mega-Scale Vision Dataset Continually with... code Syntology ran 3 of 3 samples · 0 unverified Compare
TCMP-300 (11 rows) Swin-Base — — — Compare
CINIC-10 (9 rows) VIT-L/16 (Spinal FC, Background) Reduction of Class Activation Uncertainty with Background Information code — Compare
Clothing1M (using clean data) (9 rows) CurriculumNet CurriculumNet: Weakly Supervised Learning from Large-Scale Web Images code — Compare
GasHisSDB (8 rows) CoAtNet-1 CoAtNet: Marrying Convolution and Attention for All Data Sizes code Syntology ran 2 of 5 samples · 3 unverified Compare
EMNIST-Digits (7 rows) WaveMixLite-112/16 WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
Places365 (7 rows) OmniVec2 OmniVec2 - A Novel Transformer based Network for Large Scale... — — Compare
smallNORB (7 rows) Heinsen Routing An Algorithm for Routing Capsules in All Domains code — Compare
Tiered ImageNet 5-way (5-shot) (7 rows) EGNN+Transduction Edge-labeling Graph Neural Network for Few-shot Learning code Syntology ran 0 of 1 samples · 1 unverified Compare
Colored-MNIST(with spurious correlation) (6 rows) MLP-DecAug DecAug: Out-of-Distribution Generalization via Decomposed Feature... code — Compare
iWildCam2020-WILDS (6 rows) COSMO Reviving the Context: Camera Trap Species Classification as Link... code — Compare
Oxford-IIIT Pets (6 rows) CeiT-S (384 finetune resolution) Incorporating Convolution Designs into Visual Transformers code Syntology ran 9 of 11 samples · 2 unverified Compare
Caltech-256 (5 rows) AG-Net Attend and Guide (AG-Net): A Keypoints-driven Attention-based Deep... code — Compare
Oxford-IIIT Pet Dataset (5 rows) TWIST (ResNet-50) Self-Supervised Learning by Estimating Twin Class Distributions code Syntology ran 5 of 15 samples · 10 unverified Compare
Red MiniImageNet 20% label noise (5 rows) NCR (ResNet-18) Learning with Neighbor Consistency for Noisy Labels code Syntology ran 4 of 4 samples · 0 unverified Compare
Red MiniImageNet 40% label noise (5 rows) NCR (ResNet-18) Learning with Neighbor Consistency for Noisy Labels code Syntology ran 4 of 4 samples · 0 unverified Compare
Red MiniImageNet 80% label noise (5 rows) NCR (ResNet-18) Learning with Neighbor Consistency for Noisy Labels code Syntology ran 4 of 4 samples · 0 unverified Compare
CIFAR-10 (with noisy labels) (4 rows) SSR SSR: An Efficient and Robust Framework for Learning with Unknown... code — Compare
CUB (4 rows) Entropy-based Logic Explained Network Entropy-based Logic Explanations of Neural Networks code Syntology ran 1 of 6 samples · 5 unverified Compare
Food-101N (4 rows) LRA-diffusion (CLIP ViT) Label-Retrieval-Augmented Diffusion Models for Learning from Noisy Labels code Syntology ran 7 of 9 samples · 2 unverified Compare
ISIC2018 (4 rows) UniNet UniNet: A Contrastive Learning-guided Unified Framework with... code — Compare
JFT-300M (4 rows) V-MoE-H/14 (Every-2) Scaling Vision with Sparse Mixture of Experts code Syntology ran 1 of 1 samples · 0 unverified Compare
MAMe (4 rows) EfficientNet-B3 The MAMe Dataset: On the relevance of High Resolution and Variable... code — Compare
N-MNIST (4 rows) STS-ResNet Convolutional Spiking Neural Networks for Spatio-Temporal Feature... code — Compare
ObjectNet (Bounding Box) (4 rows) BiT-L (ResNet) Big Transfer (BiT): General Visual Representation Learning code Syntology ran 3 of 10 samples · 7 unverified Compare
Oracle-MNIST (4 rows) ResNet-18 + Vision Eagle Attention Vision Eagle Attention: a new lens for advancing image classification code — Compare
Places365-Standard (4 rows) SWAG (ViT H/14) Revisiting Weakly Supervised Pre-Training of Visual Perception Models code — Compare
Red MiniImageNet 60% label noise (4 rows) InstanceGM-SS Instance-Dependent Noisy Label Learning via Graphical Modelling code — Compare
Tiny-ImageNet (4 rows) UPANets UPANets: Learning from the Universal Pixel Attention Networks code — Compare
Visual Wake Words (4 rows) HyT-NAS-BA HyT-NAS: Hybrid Transformers Neural Architecture Search for Edge Devices — — Compare
BreakHis (3 rows) WaveMix Which Backbone to Use: A Resource-efficient Domain Specific... code — Compare
CelebA 64x64 (3 rows) cFlow Null-sampling for Interpretable and Fair Representations code — Compare
EarlyNSD (3 rows) DenseNet121_256x256_Nutrispace Nutrispace: A novel color space to enhance deep learning based... — — Compare
EuroSAT-SAR (3 rows) FG-MAE (ViT-S/16) Feature Guided Masked Autoencoder for Self-supervised Learning in... code — Compare
FlickrLogos-32 (3 rows) TC-VII (with outside data) Deep Learning for Logo Recognition — — Compare
Id Pattern Dataset (3 rows) Claude 3 Opus Identification of Stone Deterioration Patterns with Large Multimodal Models code — Compare
Kvasir (3 rows) HiFuse_Small HiFuse: Hierarchical Multi-Scale Feature Fusion Network for... code — Compare
Malaria Dataset (3 rows) kEffNet-B0 V2 16ch An Enhanced Scheme for Reducing the Complexity of Pointwise... code — Compare
N-Caltech 101 (3 rows) mMND (STDP) Sequence Approximation using Feedforward Spiking Neural Network... — — Compare
SIPaKMeD (3 rows) DL+PCA+GWO Cervical Cytology Classification Using PCA & GWO Enhanced Deep... code — Compare
Causal3DIdent (2 rows) SimCLR Self-Supervised Learning with Data Augmentations Provably Isolates... code Syntology ran 3 of 3 samples · 0 unverified Compare
Certificate Verification (2 rows) ResMLP-24 ResMLP: Feedforward networks for image classification with... code Syntology ran 2 of 7 samples · 5 unverified Compare
CIFAR-10 (40 Labels, ImageNet-100 Unlabeled) (2 rows) UnMixMatch Scaling Up Semi-supervised Learning with Unconstrained Unlabelled Data code — Compare
CIFAR-10, 40% Symmetric Noise (2 rows) FaMUS Faster Meta Update Strategy for Noise-Robust Deep Learning code Syntology ran 2 of 7 samples · 5 unverified Compare
CIFAR-10, 60% Symmetric Noise (2 rows) MentorMix Faster Meta Update Strategy for Noise-Robust Deep Learning code Syntology ran 2 of 7 samples · 5 unverified Compare
CIFAR-10 Image Classification (2 rows) ASF-former-S Adaptive Split-Fusion Transformer code — Compare
CIFAR-100, 40% Symmetric Noise (2 rows) FaMUS Faster Meta Update Strategy for Noise-Robust Deep Learning code Syntology ran 2 of 7 samples · 5 unverified Compare
CLEVR/Count (2 rows) SEER (RegNet10B) Vision Models Are More Robust And Fair When Pretrained On... code — Compare
CLEVR/Dist (2 rows) SEER (RegNet10B) Vision Models Are More Robust And Fair When Pretrained On... code — Compare
CUB-200-2011 (2 rows) Sparse-CBM Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning code — Compare
Fracture/Normal Shoulder Bone X-ray Images on MURA (2 rows) Our Ensemble Learning-2 Classification of Shoulder X-Ray Images with Deep Learning Ensemble Models — — Compare
Galaxy10 DECals (2 rows) WaveMix WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
Gaze-CIFAR-10 (2 rows) DSGE-ConvNeXtV2 Gaze-Guided Learning: Avoiding Shortcut Bias in Visual Classification code — Compare
HErlev (2 rows) Fuzzy Distance Ensemble A fuzzy distance-based ensemble of deep models for cervical cancer... code — Compare
ImageNet-10 (2 rows) ResNet-50 + UDA+AutoDropout AutoDropout: Learning Dropout Patterns to Regularize Deep Networks code — Compare
ImageNet-100 (TEMI Split) (2 rows) SparseSwin with L2 SparseSwin: Swin Transformer with Sparse Transformer Block code — Compare
ImageNet-Hard (2 rows) EfficientNet-L2-Ns — — — Compare
Imagenette (2 rows) µ2Net+ (ViT-L/16) A Continual Development Methodology for Large-scale Multitask... code — Compare
Imbalanced CUB-200-2011 (2 rows) Multi-task A New Periocular Dataset Collected by Mobile Devices in... — — Compare
Intel Image Classification (2 rows) ResNet-18 + Vision Eagle Attention Vision Eagle Attention: a new lens for advancing image classification code — Compare
ISIC 2018 (2 rows) HiFuse_Base HiFuse: Hierarchical Multi-Scale Feature Fusion Network for... code — Compare
Large Labelled Logo Dataset (L3D) (2 rows) L3D_original_2level The Large Labelled Logo Dataset (L3D): A Multipurpose and... code — Compare
LIMUC (2 rows) Inception-v3 Class Distance Weighted Cross-Entropy Loss for Ulcerative Colitis... code Syntology ran 0 of 7 samples · 7 unverified Compare
Noisy MNIST (AWGN) (2 rows) PCGAN-CHAR PCGAN-CHAR: Progressively Trained Classifier Generative... — — Compare
Noisy MNIST (Contrast) (2 rows) PCGAN-CHAR PCGAN-CHAR: Progressively Trained Classifier Generative... — — Compare
Noisy MNIST (Motion) (2 rows) PCGAN-CHAR PCGAN-CHAR: Progressively Trained Classifier Generative... — — Compare
ObjectNet (ImageNet classes) (2 rows) Diffusion Classifier (zero-shot) Your Diffusion Model is Secretly a Zero-Shot Classifier code Syntology ran 2 of 2 samples · 0 unverified Compare
PlantDoc (2 rows) SCOLD A Vision-Language Foundation Model for Leaf Disease Identification code — Compare
PlantVillage (2 rows) adaptive minimal ensembling Improving plant disease classification by adaptive minimal ensembling — — Compare
split CIFAR-100 (2 rows) OFSCIL 12 mJ per Class On-Device Online Few-Shot Class-Incremental Learning code — Compare
WebVision (2 rows) PropMix (Ours) PropMix: Hard Sample Filtering and Proportional MixUp for Learning... code — Compare
AIDER (1 row) TakuNet FP=16 TakuNet: an Energy-Efficient CNN for Real-Time Inference on... code — Compare
AIDERV2 (1 row) TakuNet FP=16 TakuNet: an Energy-Efficient CNN for Real-Time Inference on... code — Compare
AmsterTime (1 row) AP-GeM (ResNet-101) AmsterTime: A Visual Place Recognition Benchmark Dataset for... code — Compare
ArtDL (1 row) ResNet-50 A Data Set and a Convolutional Model for Iconography... code — Compare
CARS196 (1 row) µ2Net+ (ViT-L/16) A Continual Development Methodology for Large-scale Multitask... code — Compare
cats_vs_dogs (1 row) µ2Net+ (ViT-L/16) A Continual Development Methodology for Large-scale Multitask... code — Compare
Chaoyang (1 row) HSANR Hard Sample Aware Noise Robust Learning for Histopathology Image... code — Compare
CIFAR-100, 60% Symmetric Noise (1 row) MentorMix Faster Meta Update Strategy for Noise-Robust Deep Learning code Syntology ran 2 of 7 samples · 5 unverified Compare
CIFAR-100 (alpha=0, 20 clients per round) (1 row) FedAvgM + ASAM + SWA Improving Generalization in Federated Learning by Seeking Flat Minima code Syntology ran 1 of 1 samples · 0 unverified Compare
CIFAR-100C (1 row) Astroformer Astroformer: More Data Might not be all you need for Classification code Syntology ran 1 of 1 samples · 0 unverified Compare
cifar-10,4000 (1 row) WRN-28-2 + UDA+AutoDropout AutoDropout: Learning Dropout Patterns to Regularize Deep Networks code — Compare
cifar10 (1 row) SAM — — — Compare
cifar100 (1 row) shreynet Deep Residual Learning for Image Recognition code Syntology ran 230 of 377 samples · 147 unverified Compare
Deep PCB (1 row) ResNet Improving Model Performance and Removing the Class Imbalance... — — Compare
DVS128 Gesture (1 row) SNN Sneaky Spikes: Uncovering Stealthy Backdoor Attacks in Spiking... code — Compare
EMNIST-Byclass (1 row) WaveMixLite-128/7 WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
EMNIST-Bymerge (1 row) WaveMixLite-128/16 WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
ESC-50 (1 row) SDGM-D Performance of Gaussian Mixture Model Classifiers on Embedded... code — Compare
FEMNIST (1 row) pFedBreD_ns_mg Personalized Federated Learning with Hidden Information on... — — Compare
FGVC Aircraft (1 row) TransBoost-ResNet50 TransBoost: Improving the Best ImageNet Performance using Deep Transduction code — Compare
FGVC-Aircraft (1 row) EnGraf-Net101 (G=4, H=1) EnGraf-Net: Multiple Granularity Branch Network with Fine-Coarse... code — Compare
Flower102 (1 row) efficient adaptive ensembling Efficient Adaptive Ensembling for Image Classification — — Compare
Flowers (Tensorflow) (1 row) CNN+ Wilson-Cowan model RNN Learning in Wilson-Cowan model for metapopulation code — Compare
FMD (materials) (1 row) RADAM (ConvNeXt-L) RADAM: Texture Recognition through Randomized Aggregated Encoding... code — Compare
GTSRB (1 row) SAG-ViT SAG-ViT: A Scale-Aware, High-Fidelity Patching Approach with Graph... code — Compare
iCassava'19 (1 row) E2E-3M Rethinking Recurrent Neural Networks and Other Improvements for... code — Compare
delete (1 row) (unnamed in the archive) Understanding the Robustness of Randomized Feature Defense Against... code Syntology ran 15 of 20 samples · 5 unverified Compare
ImageNet-100 (Class-IL, 5T) (1 row) MoCo + CaSSLe Regularizing with Pseudo-Negatives for Continual Self-Supervised Learning code Syntology ran 1 of 1 samples · 0 unverified Compare
imagenet-1k (1 row) BinaryViT BinaryViT: Pushing Binary Vision Transformers Towards Convolutional Models code — Compare
ImageNet-32 (1 row) WRN (N=28, k=10) A Downsampled Variant of ImageNet as an Alternative to the CIFAR datasets code — Compare
ImageNet-64 (1 row) WRN (N=36, k=5) A Downsampled Variant of ImageNet as an Alternative to the CIFAR datasets code — Compare
ImageNet-9 (1 row) SqueezeNet + Simple Bypass SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and... code Syntology ran 4 of 4 samples · 0 unverified Compare
ImageNet-P (1 row) SqueezeNet + Simple Bypass SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and... code Syntology ran 4 of 4 samples · 0 unverified Compare
ImageNet-Sketch (1 row) µ2Net+ (ViT-L/16) A Continual Development Methodology for Large-scale Multitask... code — Compare
iNat2021-mini (1 row) WaveMix-256/16 (level 2) WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
ISBNet (1 row) ThanosNet ThanosNet: A Novel Trash Classification Method Using Metadata code — Compare
KITTI-Dist (1 row) SEER (RegNet10B) Vision Models Are More Robust And Fair When Pretrained On... code — Compare
KMNIST (1 row) µ2Net (ViT-L/16) An Evolutionary Approach to Dynamic Introduction of Tasks in... code — Compare
KTH-TIPS2 (1 row) RADAM (ConvNeXt-XL) RADAM: Texture Recognition through Randomized Aggregated Encoding... code — Compare
LabelMe (1 row) CoNAL Learning from Crowds by Modeling Common Confusions code — Compare
LeafNet (1 row) SCOLD A Vision-Language Foundation Model for Leaf Disease Identification code — Compare
mnist (1 row) WaveMixLite WaveMix: A Resource-efficient Neural Network for Image Analysis code — Compare
MNIST-rot-12 (1 row) PDO-eConv (ours) PDO-eConvs: Partial Differential Operator Based Equivariant Convolutions code Syntology ran 0 of 7 samples · 7 unverified Compare
MNIST-rot-12k (DA) (1 row) PDO-eConv (ours) PDO-eConvs: Partial Differential Operator Based Equivariant Convolutions code Syntology ran 0 of 7 samples · 7 unverified Compare
MultiMNIST (1 row) CapsNet Dynamic Routing Between Capsules code Syntology ran 18 of 118 samples · 100 unverified Compare
NCT-CRC-HE-100K (1 row) SAG-ViT SAG-ViT: A Scale-Aware, High-Fidelity Patching Approach with Graph... code — Compare
No Background RGB Arabic Alphabets Sign Language Dataset (1 row) ArabSignNet Deep Learning Recognition for Arabic Alphabet Sign Language RGB Dataset — — Compare
PASCAL VOC 2007 (1 row) NNCLR With a Little Help from My Friends: Nearest-Neighbor Contrastive... code Syntology ran 4 of 5 samples · 1 unverified Compare
Pets SAM (1 row) efficient adaptive ensembling Efficient Adaptive Ensembling for Image Classification — — Compare
PRImA (1 row) ResNet-152 2x (RS training) Revisiting ResNets: Improved Training and Scaling Strategies code — Compare
QMNIST (1 row) Deep regularization Deep regularization and direct training of the inner layers of... code — Compare
RGB Arabic Alphabet Sign Language (AASL) dataset (1 row) ArabSignNet Deep Learning Recognition for Arabic Alphabet Sign Language RGB Dataset — — Compare
SARS-COV-2 (1 row) Fuzzy rank-based fusion of CNN models using Gompertz function Fuzzy Rank-based Fusion of CNN Models using Gompertz Function for... code — Compare
So2Sat LCZ42 (1 row) ResNet50 In-domain representation learning for remote sensing code — Compare
Split CIFAR-10 (1 row) Model with negotiation paradigm Negotiated Representations to Prevent Forgetting in Machine... code — Compare
Split Fashion M-NIST (1 row) Model with negotiation paradigm Negotiated Representations to Prevent Forgetting in Machine... code — Compare
Split M-NIST (1 row) Model with negotiation paradigm Negotiated Representations to Prevent Forgetting in Machine... code — Compare
Sports10 (1 row) Max Margin Contrastive Contrastive Learning of Generalized Game Representations code — Compare
Stanford Online Products (1 row) µ2Net+ (ViT-L/16) A Continual Development Methodology for Large-scale Multitask... code — Compare
SUN397 (1 row) TransBoost-ResNet50 TransBoost: Improving the Best ImageNet Performance using Deep Transduction code — Compare
Surrey ASL (1 row) E2E-3M Rethinking Recurrent Neural Networks and Other Improvements for... code — Compare
Training and validation dataset of capsule vision 2024 challenge. (1 row) BiomedCLIP+PubmedBERT A Multimodal Approach For Endoscopic VCE Image Classification... code — Compare
VizWiz-Classification (1 row) VOLO-D5 VOLO: Vision Outlooker for Visual Recognition code Syntology ran 1 of 6 samples · 5 unverified Compare
chbh7051/vit-base-driver-drowsiness-detection (0 rows) no rows in the archive — —
Custom Dataset (0 rows) no rows in the archive — —
Custom DeepFake Dataset (0 rows) no rows in the archive — —
finetuned-websites (0 rows) no rows in the archive — —
Human_Action_Recognition (0 rows) no rows in the archive — —
image_folder (0 rows) no rows in the archive — —
imagefolder (0 rows) no rows in the archive — —
indian_food_images (0 rows) no rows in the archive — —
mriDataSet (0 rows) no rows in the archive — —
New Plant Diseases Dataset (0 rows) no rows in the archive — —
PCam (0 rows) no rows in the archive — —

Syntology column: samples harvested from the paper's repositories and executed on synthesized fixtures; “ran” is not a correctness claim and does not order the table. A dash means no Syntology record for that paper, not a recorded non-run. Read from the graph 2026-09-24.

Libraries

Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.

Datasets archive 2025-07-28

283 datasets whose archive record lists this task, ordered by the archive's paper count. 30 shown of 283 until expanded.

CIFAR-10ImageNetCIFAR-100MNISTCelebASVHNFashion-MNISTCUB-200-2011Oxford 102 FlowerTiny ImageNetSTL-10DTDFood-101Stanford CarsEuroSATiNaturalistPlaces205FGVC-AircraftBDD100KCaltech-256ESC-50GTSRBtieredImageNetClothing1MImageNet-SketchEMNISTNAS-Bench-201YFCC100MStanford Online ProductsVGG-SoundAI2DVTABCINIC-10RESISC45Extended Yale BWebVisionLabelMeObjectNetfMoWOxford5kMeta-DatasetPASCAL VOC 2007JFT-300MKvasirImageNet-32smallNORBN-Caltech 101PCamStylized ImageNetDVS128 GestureTiny ImagesCIFAR-10NKuzushiji-MNISTOxford-IIIT Pet DatasetImageNet-OAbstractReasoningBigEarthNetChestX-ray8PlantVillageCIFAR-100NPlaces365Oxford-IIIT PetsWOSPGMMINCTiny-ImageNet-CMultiMNISTSUN397CARS196JFT-3BMillion-AIDVisual Wake WordsImagenetteSUN AttributeOpen Images V4ImageNet-64ImageNet-POmniBenchmarkEyeQiNat2021FlickrLogos-32QMNISTUrbanCarsBAM!ELEVATERImageNet-WIP102VizWiz-ClassificationLAGMSegETHOSPlantDocChaoyangDeepFishLaMemMLRSNetFEMNISTFood2KN-MNISTTAU Urban Acoustic Scenes 2019Cats and DogsCausal3DIdentHyper-Kvasir DatasetImageNet-100 (TEMI Split)BCN_20000BCNBBreakHisFoodX-251GasHisSDBSo2Sat LCZ42Kuzushiji-49AmsterTimeiCartoonFaceiWildCam2020-WILDSKMNISTAIDERColonINST-v1ColonINST-v1 (Seen)ColonINST-v1 (Unseen)NCT-CRC-HE-100KArtBench-10 (32x32)BambooBook Cover DatasetDF20DFUC2021Grocery StoreImageNet-9Kannada-MNISTKaoKoreNumtaDBPMDataPS-BattlesSI-SCOREColored-MNIST(with spurious correlation)F-CelebA (10 tasks)Food-101NGalaxy Zoo DECaLSMalaria DatasetOOD-CVOpen Images V7Red MiniImageNet 20% label noiseRed MiniImageNet 40% label noiseRed MiniImageNet 80% label noiseUrban EnvironmentsHErlevImageNet-HardImageNet-PatchInsPLADIntel Image ClassificationMuMiNNIH-CXR-LTRF100Sewer-MLSIPaKMeDSports10Tencent ML-ImagesCI-MNISTDiagSetETHECEuroSAT-SARKTH-TIPS2Kuzushiji-KanjiLIMUCMIMIC-CXR-LTMNIST Large Scale datasetOracle-MNISTWHU-HiACL-FigAnimals-10AtlasDEIC BenchmarkDiffusion DeepfakeDirty-MNISTFireRiskFMD (materials)Image and Video AdvertisementsLKSLSA16OFDIWRGB Arabic Alphabets Sign Language DatasetSIDD-ImageStream-51AdvNetASIRRACARBENCHAMMICross-View Time DatasetDeep PCBDF20 - MiniDIB-10KEndotect Polyp Segmentation Challenge DatasetHiAMLInceptionInVar-100Kvasir-CapsuleMAMeNew Plant Diseases DatasetPolSFS2RDASI-ScoreSTIRTwo-PathUltra Fine-Grained Leaves (Cotton, SoyAgeing, SoyGene, SoyGlobal, SoyLocal)Vistas-NPAIDERV2AjwaOrMedjoolARC-100ArtDLBanappleBankNote-NetCervix93 Cytology DatasetCleanSTL-10CNFOOD-241CNFOOD-241-ChenCross-View Time Dataset (Cross-Camera Split)Cultural Events Classification using Hyper-parameter Optimization of Deep Learning TechniqueDalleStreetDry Bean DatasetEarlyNSDENSegFathomNet2023Fruits Dataset for ClassificationGaze-CIFAR-10HaSPeRIcon645Id Pattern DatasetidspritesImageNet 50 samples per classIMPACT PatentiNaturalist Fine-Grained GeolocationIndian Party Symbol DatasetIran's Built Heritage Binary Image Classification DatasetIRMAISBNetJAMBOLabelsLAOFIW DatasetLarge Labelled Logo Dataset (L3D)LeafNetMapReader DataMuMiN-largeMuMiN-mediumMuMiN-smallNo Background RGB Arabic Alphabets Sign Language DatasetNotre-Dame Cathedral FireOLID IOmni-ImageOpenStreetMap Multi-Sensor Scene ClassificationOrchid2024PhotozillaPoTATO: A Dataset for Analyzing Polarimetric Traces of Afloat Trash ObjectsPRImARGB Arabic Alphabet Sign Language (AASL) datasetShipSpottingSolarDKSPOT-10SSBI DatasetSuSy DatasetSVLDSynthetic COVID-19 CXR DatasetTCB-DSTEM nanowire morphologies for classification and segmentationtopex-printerTsinghua DogsTwinSynthsWeapon Detection DatasetYFCC100M Fine-Grained GeolocationADFIAlpaca Dataset Image ClassificationAppleScabFDsAppleScabLDsCCICCEAHB2021-5Corn Kernel Images DatasetCorn Seeds DataseteAppleScabFour ShapesLusitano Fabric Defect Detection DatasetMoroccan Monay datasetMudestredaPortuguese Meals DatasetSakha-TBTCMP-300

Subtasks archive 2025-07-28

33 subtasks in the archive's task tree. 30 shown of 33 until expanded.

Most implemented papers archive 2025-07-28

30 shown of 4,702 papers with code (10,488 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.

  • 10 Dec 2015 484 repositories listed Syntology ran 230 of 377 samples · 147 unverified · 187 pointer-only (licence)
    Deep residual nets are foundations of our submissions to ILSVRC & COCO 2015 competitions, where we also won the 1st places on the tasks of ImageNet detection, ImageNet localization, COCO detection, and COCO segmentation.
  • 4 Sep 2014 305 repositories listed Syntology ran 12 of 122 samples · 110 unverified · 4 pointer-only (licence)
    In this work we investigate the effect of the convolutional network depth on its accuracy in the large-scale image recognition setting.
  • 13 Jan 2018 159 repositories listed Syntology ran 85 of 111 samples · 26 unverified · 64 pointer-only (licence)
    In this paper we describe a new mobile architecture, MobileNetV2, that improves the state of the art performance of mobile models on multiple tasks and benchmarks as well as across a spectrum of different model sizes.
  • 22 Oct 2020 158 repositories listed Syntology ran 281 of 419 samples · 138 unverified · 154 pointer-only (licence)
    While the Transformer architecture has become the de-facto standard for natural language processing tasks, its applications to computer vision remain limited.
  • 25 Aug 2016 146 repositories listed Syntology ran 18 of 71 samples · 53 unverified · 7 pointer-only (licence)
    Recent work has shown that convolutional networks can be substantially deeper, more accurate, and efficient to train if they contain shorter connections between layers close to the input and those close to the output.
  • 28 May 2019 144 repositories listed Syntology ran 171 of 302 samples · 131 unverified · 112 pointer-only (licence)
    Convolutional Neural Networks (ConvNets) are commonly developed at a fixed resource budget, and then scaled up for better accuracy if more resources are available.
  • 7 Oct 2016 126 repositories listed Syntology ran 79 of 141 samples · 62 unverified · 68 pointer-only (licence)
    For captioning and VQA, we show that even non-attention based models can localize inputs.
  • 7 Oct 2016 126 repositories listed Syntology ran 79 of 141 samples · 62 unverified · 68 pointer-only (licence)
    For captioning and VQA, we show that even non-attention based models can localize inputs.
  • 27 Nov 2019 123 repositories listed
    Neural networks have enabled state-of-the-art approaches to achieve incredible results on computer vision tasks such as object detection.
  • 2 Dec 2015 113 repositories listed Syntology ran 5 of 26 samples · 21 unverified · 4 pointer-only (licence)
    Convolutional networks are at the core of most state-of-the-art computer vision solutions for a wide variety of tasks.
  • 13 Feb 2020 96 repositories listed Syntology ran 79 of 137 samples · 58 unverified · 52 pointer-only (licence)
    This paper presents SimCLR: a simple framework for contrastive learning of visual representations.
  • 23 Feb 2016 87 repositories listed Syntology ran 1 of 1 samples · 0 unverified
    Recently, the introduction of residual connections in conjunction with a more traditional architecture has yielded state-of-the-art performance in the 2015 ILSVRC challenge; its performance was similar to the latest…
  • 5 Sep 2017 85 repositories listed Syntology ran 2 of 9 samples · 7 unverified · 6 pointer-only (licence)
    Squeeze-and-Excitation Networks formed the foundation of our ILSVRC 2017 classification submission which won first place and reduced the top-5 error to 2.
  • 9 Mar 2017 85 repositories listed Syntology ran 86 of 154 samples · 68 unverified · 57 pointer-only (licence)
    We propose an algorithm for meta-learning that is model-agnostic, in the sense that it is compatible with any model trained with gradient descent and applicable to a variety of different learning problems, including…
  • 9 Mar 2017 85 repositories listed Syntology ran 86 of 154 samples · 68 unverified · 57 pointer-only (licence)
    We propose an algorithm for meta-learning that is model-agnostic, in the sense that it is compatible with any model trained with gradient descent and applicable to a variety of different learning problems, including…
  • 17 Sep 2014 83 repositories listed Syntology ran 27 of 42 samples · 15 unverified · 20 pointer-only (licence)
    We propose a deep convolutional neural network architecture codenamed "Inception", which was responsible for setting the new state of the art for classification and detection in the ImageNet Large-Scale Visual…
  • 26 Feb 2021 82 repositories listed Syntology ran 16 of 20 samples · 4 unverified · 16 pointer-only (licence)
    State-of-the-art computer vision systems are trained to predict a fixed set of predetermined object categories.
  • 25 Mar 2021 80 repositories listed Syntology ran 108 of 207 samples · 99 unverified · 43 pointer-only (licence)
    This paper presents a new vision Transformer, called Swin Transformer, that capably serves as a general-purpose backbone for computer vision.
  • 25 Mar 2021 80 repositories listed Syntology ran 108 of 207 samples · 99 unverified · 43 pointer-only (licence)
    This paper presents a new vision Transformer, called Swin Transformer, that capably serves as a general-purpose backbone for computer vision.
  • 7 Feb 2018 78 repositories listed Syntology ran 43 of 72 samples · 29 unverified · 40 pointer-only (licence)
    The former networks are able to encode multi-scale contextual information by probing the incoming features with filters or pooling operations at multiple rates and multiple effective fields-of-view, while the latter…
  • 26 Oct 2017 77 repositories listed Syntology ran 18 of 118 samples · 100 unverified · 12 pointer-only (licence)
    We use the length of the activity vector to represent the probability that the entity exists and its orientation to represent the instantiation parameters.
  • 23 May 2016 72 repositories listed Syntology ran 60 of 96 samples · 36 unverified · 46 pointer-only (licence)
    Deep residual networks were shown to be able to scale up to thousands of layers and still have improving performance.
  • 25 Oct 2017 71 repositories listed Syntology ran 30 of 47 samples · 17 unverified · 15 pointer-only (licence)
    We also find that mixup reduces the memorization of corrupt labels, increases the robustness to adversarial examples, and stabilizes the training of generative adversarial networks.
  • 11 Feb 2015 70 repositories listed Syntology ran 14 of 21 samples · 7 unverified · 5 pointer-only (licence)
    Training Deep Neural Networks is complicated by the fact that the distribution of each layer's inputs changes during training, as the parameters of the previous layers change.
  • 11 Feb 2015 70 repositories listed Syntology ran 14 of 21 samples · 7 unverified · 5 pointer-only (licence)
    Training Deep Neural Networks is complicated by the fact that the distribution of each layer's inputs changes during training, as the parameters of the previous layers change.
  • 6 May 2019 67 repositories listed Syntology ran 58 of 105 samples · 47 unverified · 46 pointer-only (licence)
    We achieve new state of the art results for mobile classification, detection and segmentation.
  • 4 Dec 2016 67 repositories listed Syntology ran 7 of 29 samples · 22 unverified · 5 pointer-only (licence)
    Scene parsing is challenging for unrestricted open vocabulary and diverse scenes.
  • 16 Nov 2016 61 repositories listed Syntology ran 34 of 80 samples · 46 unverified · 13 pointer-only (licence)
    Our simple design results in a homogeneous, multi-branch architecture that has only a few hyper-parameters to set.
  • 16 Nov 2016 61 repositories listed Syntology ran 34 of 80 samples · 46 unverified · 13 pointer-only (licence)
    Our simple design results in a homogeneous, multi-branch architecture that has only a few hyper-parameters to set.
  • 24 Jun 2018 59 repositories listed Syntology ran 66 of 156 samples · 90 unverified · 48 pointer-only (licence)
    This paper addresses the scalability challenge of architecture search by formulating the task in a differentiable manner.

Syntology lines on 29 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections