{"url":"/dataset/imagenet","name":"ImageNet","full_name":null,"description_markdown":"The **ImageNet** dataset contains 14,197,122 annotated images according to the WordNet hierarchy. Since 2010 the dataset is used in the ImageNet Large Scale Visual Recognition Challenge (ILSVRC), a benchmark in image classification and object detection.\r\nThe publicly released dataset contains a set of manually annotated training images. A set of test images is also released, with the manual annotations withheld.\r\nILSVRC annotations fall into one of two categories: (1) image-level annotation of a binary label for the presence or absence of an object class in the image, e.g., “there are cars in this image” but “there are no tigers,” and (2) object-level annotation of a tight bounding box and class label around an object instance in the image, e.g., “there is a screwdriver centered at position (20,25) with width of 50 pixels and height of 30 pixels”.\r\nThe ImageNet project does not own the copyright of the images, therefore only thumbnails and URLs of images are provided.\r\n\r\n* Total number of non-empty WordNet synsets: 21841\r\n* Total number of images: 14197122\r\n* Number of images with bounding box annotations: 1,034,908\r\n* Number of synsets with SIFT features: 1000\r\n* Number of images with SIFT features: 1.2 million\r\n\r\nSource: [ImageNet Large Scale Visual Recognition Challenge](https://arxiv.org/abs/1409.0575)\r\nImage Source: [https://cs.stanford.edu/people/karpathy/cnnembed/](https://cs.stanford.edu/people/karpathy/cnnembed/)","description_withheld":null,"homepage":"https://image-net.org/index.php","introduced_date":"2009-01-01","introduced_date_note":null,"introduced_by":{"paper":null,"title":"ImageNet: A large-scale hierarchical image database","first_author":null,"url":"https://doi.org/10.1109/CVPR.2009.5206848"},"license":{"name":"Custom (research, non-commercial)","url":"http://image-net.org/"},"modalities":[{"name":"Images","url":"/datasets/modality/images"}],"tasks":[{"name":"Image Classification","url":"/task/image-classification","datasets_with_task":"/datasets/task/image-classification"},{"name":"Classification","url":"/task/classification-1","datasets_with_task":"/datasets/task/classification-1"},{"name":"Image Generation","url":"/task/image-generation","datasets_with_task":"/datasets/task/image-generation"},{"name":"Zero-Shot Learning","url":"/task/zero-shot-learning","datasets_with_task":"/datasets/task/zero-shot-learning"},{"name":"Visual Question Answering (VQA)","url":"/task/visual-question-answering","datasets_with_task":"/datasets/task/visual-question-answering"},{"name":"Few-Shot Image Classification","url":"/task/few-shot-image-classification","datasets_with_task":"/datasets/task/few-shot-image-classification"},{"name":"Image Super-Resolution","url":"/task/image-super-resolution","datasets_with_task":"/datasets/task/image-super-resolution"},{"name":"Color Image Denoising","url":"/task/color-image-denoising","datasets_with_task":"/datasets/task/color-image-denoising"},{"name":"Few-Shot Learning","url":"/task/few-shot-learning","datasets_with_task":"/datasets/task/few-shot-learning"},{"name":"Semi-Supervised Image Classification","url":"/task/semi-supervised-image-classification","datasets_with_task":"/datasets/task/semi-supervised-image-classification"},{"name":"Image Clustering","url":"/task/image-clustering","datasets_with_task":"/datasets/task/image-clustering"},{"name":"Image Reconstruction","url":"/task/image-reconstruction","datasets_with_task":"/datasets/task/image-reconstruction"},{"name":"Neural Architecture Search","url":"/task/architecture-search","datasets_with_task":"/datasets/task/architecture-search"},{"name":"Transductive Zero-Shot Classification","url":"/task/transductive-zero-shot-classification","datasets_with_task":"/datasets/task/transductive-zero-shot-classification"},{"name":"Weakly Supervised Object Detection","url":"/task/weakly-supervised-object-detection","datasets_with_task":"/datasets/task/weakly-supervised-object-detection"},{"name":"Image Inpainting","url":"/task/image-inpainting","datasets_with_task":"/datasets/task/image-inpainting"},{"name":"Binarization","url":"/task/binarization","datasets_with_task":"/datasets/task/binarization"},{"name":"Zero-Shot Transfer Image Classification","url":"/task/zero-shot-transfer-image-classification","datasets_with_task":"/datasets/task/zero-shot-transfer-image-classification"},{"name":"Prompt Engineering","url":"/task/prompt-engineering","datasets_with_task":"/datasets/task/prompt-engineering"},{"name":"Object Recognition","url":"/task/object-recognition","datasets_with_task":"/datasets/task/object-recognition"},{"name":"Image Segmentation","url":"/task/image-segmentation","datasets_with_task":"/datasets/task/image-segmentation"},{"name":"Adversarial Defense","url":"/task/adversarial-defense","datasets_with_task":"/datasets/task/adversarial-defense"},{"name":"Quantization","url":"/task/quantization","datasets_with_task":"/datasets/task/quantization"},{"name":"Medical Image Classification","url":"/task/medical-image-classification","datasets_with_task":"/datasets/task/medical-image-classification"},{"name":"Zero-Shot Composed Image Retrieval (ZS-CIR)","url":"/task/zero-shot-composed-image-retrieval-zs-cir","datasets_with_task":"/datasets/task/zero-shot-composed-image-retrieval-zs-cir"},{"name":"Knowledge Distillation","url":"/task/knowledge-distillation","datasets_with_task":"/datasets/task/knowledge-distillation"},{"name":"Image Deblurring","url":"/task/image-deblurring","datasets_with_task":"/datasets/task/image-deblurring"},{"name":"Weakly-Supervised Object Localization","url":"/task/weakly-supervised-object-localization","datasets_with_task":"/datasets/task/weakly-supervised-object-localization"},{"name":"Unsupervised Image Classification","url":"/task/unsupervised-image-classification","datasets_with_task":"/datasets/task/unsupervised-image-classification"},{"name":"Adversarial Robustness","url":"/task/adversarial-robustness","datasets_with_task":"/datasets/task/adversarial-robustness"},{"name":"Network Pruning","url":"/task/network-pruning","datasets_with_task":"/datasets/task/network-pruning"},{"name":"Image Compressed Sensing","url":"/task/image-compressed-sensing","datasets_with_task":"/datasets/task/image-compressed-sensing"},{"name":"Data Augmentation","url":"/task/data-augmentation","datasets_with_task":"/datasets/task/data-augmentation"},{"name":"Contrastive Learning","url":"/task/contrastive-learning","datasets_with_task":"/datasets/task/contrastive-learning"},{"name":"Model Compression","url":"/task/model-compression","datasets_with_task":"/datasets/task/model-compression"},{"name":"Sparse Learning","url":"/task/sparse-learning","datasets_with_task":"/datasets/task/sparse-learning"},{"name":"Self-Supervised Image Classification","url":"/task/self-supervised-image-classification","datasets_with_task":"/datasets/task/self-supervised-image-classification"},{"name":"Classification with Binary Neural Network","url":"/task/classification-with-binary-neural-network","datasets_with_task":"/datasets/task/classification-with-binary-neural-network"},{"name":"Image Colorization","url":"/task/image-colorization","datasets_with_task":"/datasets/task/image-colorization"},{"name":"JPEG Decompression","url":"/task/jpeg-decompression","datasets_with_task":"/datasets/task/jpeg-decompression"},{"name":"Image Classification with Differential Privacy","url":"/task/image-classification-with-dp","datasets_with_task":"/datasets/task/image-classification-with-dp"},{"name":"Zero-Shot Transfer Image Classification (CN)","url":"/task/zero-shot-transfer-image-classification-cn","datasets_with_task":"/datasets/task/zero-shot-transfer-image-classification-cn"},{"name":"Feature Upsampling","url":"/task/feature-upsampling","datasets_with_task":"/datasets/task/feature-upsampling"}],"languages":[{"name":"Chinese","url":"/datasets/language/chinese"}],"variants":["imagenet-1k","ImageNet sigma50","ImageNet sigma250","ImageNet sigma200","ImageNet sigma150","ImageNet sigma100","ImageNet100","ImageNet (linear)","ImageNet (finetuned)","ImageNet - 5-shot","ImageNet - 5 labeled data per class","ImageNet - 2 labeled data per class","ImageNet - 1-shot","ImageNet - 1 labeled data per class","ImageNet - 10-shot","ImageNet - 0.2% labeled data","ImageNetV2","ImageNet V2","ImageNet - 10% labeled data","ImageNet - 1% labeled data","ImageNet - 0-Shot","ImageNet"],"data_loaders":[{"repo":"https://github.com/huggingface/datasets","url":"https://huggingface.co/datasets/coallaoh/ImageNet-AB","frameworks":["tf","pytorch","jax"]},{"repo":"https://github.com/huggingface/datasets","url":"https://huggingface.co/datasets/slegroux/tiny-imagenet-200-clean","frameworks":["tf","pytorch","jax"]},{"repo":"https://github.com/huggingface/datasets","url":"https://huggingface.co/datasets/evanarlian/imagenet_1k_resized_256","frameworks":["tf","pytorch","jax"]},{"repo":"https://github.com/huggingface/datasets","url":"https://huggingface.co/datasets/imagenet-1k","frameworks":["tf","pytorch","jax"]},{"repo":"https://github.com/huggingface/datasets","url":"https://huggingface.co/datasets/Maysee/tiny-imagenet","frameworks":["tf","pytorch","jax"]},{"repo":"https://github.com/huggingface/datasets","url":"https://huggingface.co/datasets/zh-plus/tiny-imagenet","frameworks":["tf","pytorch","jax"]},{"repo":"https://github.com/pytorch/vision","url":"https://pytorch.org/vision/stable/generated/torchvision.datasets.ImageNet.html","frameworks":["pytorch"]},{"repo":"https://github.com/voxel51/fiftyone","url":"https://docs.voxel51.com/user_guide/dataset_zoo/datasets.html#imagenet-2012","frameworks":["tf","pytorch"]},{"repo":"https://github.com/activeloopai/Hub","url":"https://docs.activeloop.ai/datasets/imagenet-dataset","frameworks":["tf","pytorch"]},{"repo":"https://github.com/tensorflow/datasets","url":"https://www.tensorflow.org/datasets/catalog/imagenet2012","frameworks":["tf","jax"]},{"repo":"https://github.com/open-mmlab/mmclassification","url":"https://github.com/open-mmlab/mmclassification/blob/master/docs/getting_started.md","frameworks":["pytorch"]}],"num_papers_in_archive":15430,"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28"},"benchmarks":[{"leaderboard":"/sota/image-classification-on-imagenet","task":"Image Classification","dataset_variant":"ImageNet","rows":1060,"metrics":["Top 1 Accuracy","Number of params","GFLOPs","Hardware Burden","Top 5 Accuracy","Operations per network pass"],"first_row_in_archive_order":{"model":"CoCa (finetuned)","paper":"/paper/coca-contrastive-captioners-are-image-text","metrics":{"Number of params":"2100M","Top 1 Accuracy":"91.0%"},"code_links":[{"title":"mlfoundations/open_clip","url":"https://github.com/mlfoundations/open_clip"},{"title":"facebookresearch/multimodal","url":"https://github.com/facebookresearch/multimodal"},{"title":"lucidrains/CoCa-pytorch","url":"https://github.com/lucidrains/CoCa-pytorch"},{"title":"amitakamath/whatsup_vlms","url":"https://github.com/amitakamath/whatsup_vlms"},{"title":"amitakamath/hard_positives","url":"https://github.com/amitakamath/hard_positives"},{"title":"Chaolei98/FreeZAD","url":"https://github.com/Chaolei98/FreeZAD"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/self-supervised-image-classification-on","task":"Self-Supervised Image Classification","dataset_variant":"ImageNet","rows":144,"metrics":["Top 1 Accuracy","Top 5 Accuracy","Number of Params"],"first_row_in_archive_order":{"model":"DINOv2+reg (ViT-g/14)","paper":"/paper/vision-transformers-need-registers","metrics":{"Number of Params":"1100M","Top 1 Accuracy":"87.1"},"code_links":[{"title":"rwightman/pytorch-image-models","url":"https://github.com/rwightman/pytorch-image-models"},{"title":"facebookresearch/dinov2","url":"https://github.com/facebookresearch/dinov2"},{"title":"locuslab/massive-activations","url":"https://github.com/locuslab/massive-activations"},{"title":"borisdayma/clip-jax","url":"https://github.com/borisdayma/clip-jax"},{"title":"nickjiang2378/test-time-registers","url":"https://github.com/nickjiang2378/test-time-registers"},{"title":"birder/birder","url":"https://gitlab.com/birder/birder"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/neural-architecture-search-on-imagenet","task":"Neural Architecture Search","dataset_variant":"ImageNet","rows":135,"metrics":["Top-1 Error Rate","Accuracy","Params","MACs","FLOPs"],"first_row_in_archive_order":{"model":"DeepMAD-50M","paper":"/paper/deepmad-mathematical-architecture-design-for","metrics":{"Accuracy":"83.9","FLOPs":"8.7G","Params":"50M","Top-1 Error Rate":"16.1"},"code_links":[{"title":"alibaba/lightweight-neural-architecture-search","url":"https://github.com/alibaba/lightweight-neural-architecture-search"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/semi-supervised-image-classification-on-2","task":"Semi-Supervised Image Classification","dataset_variant":"ImageNet - 10% labeled data","rows":75,"metrics":["Top 1 Accuracy","Top 5 Accuracy","Number of params"],"first_row_in_archive_order":{"model":"DHO (ViT-Large)","paper":"/paper/simple-semi-supervised-knowledge-distillation","metrics":{"Top 1 Accuracy":"85.9%"},"code_links":[{"title":"erjui/DHO","url":"https://github.com/erjui/DHO"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/self-supervised-image-classification-on-1","task":"Self-Supervised Image Classification","dataset_variant":"ImageNet (finetuned)","rows":65,"metrics":["Top 1 Accuracy","Number of Params"],"first_row_in_archive_order":{"model":"DINOv2 (ViT-g/14, 448)","paper":"/paper/dinov2-learning-robust-visual-features","metrics":{"Number of Params":"1100M","Top 1 Accuracy":"88.9%"},"code_links":[{"title":"huggingface/transformers","url":"https://github.com/huggingface/transformers"},{"title":"facebookresearch/dinov2","url":"https://github.com/facebookresearch/dinov2"},{"title":"roboflow/rf-detr","url":"https://github.com/roboflow/rf-detr"},{"title":"open-edge-platform/training_extensions","url":"https://github.com/open-edge-platform/training_extensions"},{"title":"OML-Team/open-metric-learning","url":"https://github.com/OML-Team/open-metric-learning"},{"title":"leondgarse/keras_cv_attention_models","url":"https://github.com/leondgarse/keras_cv_attention_models/tree/main/keras_cv_attention_models/beit"},{"title":"fabio-sim/Depth-Anything-ONNX","url":"https://github.com/fabio-sim/Depth-Anything-ONNX"},{"title":"open-edge-platform/geti","url":"https://github.com/open-edge-platform/geti"},{"title":"facebookresearch/highrescanopyheight","url":"https://github.com/facebookresearch/highrescanopyheight"},{"title":"PaddlePaddle/PASSL","url":"https://github.com/PaddlePaddle/PASSL"},{"title":"beneroth13/dinov2","url":"https://github.com/beneroth13/dinov2"},{"title":"mohammedsb/dinov2formedical","url":"https://github.com/mohammedsb/dinov2formedical"},{"title":"marrlab/dinobloom","url":"https://github.com/marrlab/dinobloom"},{"title":"bespontaneous/proteus-pytorch","url":"https://github.com/bespontaneous/proteus-pytorch"},{"title":"ByungKwanLee/Causal-Unsupervised-Segmentation","url":"https://github.com/ByungKwanLee/Causal-Unsupervised-Segmentation"},{"title":"zhu-xlab/softcon","url":"https://github.com/zhu-xlab/softcon"},{"title":"birder/birder","url":"https://gitlab.com/birder/birder"},{"title":"gorkaydemir/DINOSAUR","url":"https://github.com/gorkaydemir/DINOSAUR"},{"title":"seatizendoi/dinovdeau","url":"https://github.com/seatizendoi/dinovdeau"},{"title":"BurguerJohn/global_perceptual_similarity_loss","url":"https://github.com/BurguerJohn/global_perceptual_similarity_loss"},{"title":"buyeah1109/KEN","url":"https://github.com/buyeah1109/KEN"},{"title":"JHKim-snu/PGA","url":"https://github.com/JHKim-snu/PGA"},{"title":"BurguerJohn/torch-felix","url":"https://github.com/BurguerJohn/torch-felix"},{"title":"2024-MindSpore-1/Code2","url":"https://github.com/2024-MindSpore-1/Code2/tree/main/model-1/dinov2"},{"title":"buyeah1109/finc","url":"https://github.com/buyeah1109/finc"},{"title":"pwc-1/Paper-8","url":"https://github.com/pwc-1/Paper-8/tree/main/dinov2"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/semi-supervised-image-classification-on-1","task":"Semi-Supervised Image Classification","dataset_variant":"ImageNet - 1% labeled data","rows":65,"metrics":["Top 1 Accuracy","Top 5 Accuracy","Number of params"],"first_row_in_archive_order":{"model":"DHO (ViT-Large)","paper":"/paper/simple-semi-supervised-knowledge-distillation","metrics":{"Top 1 Accuracy":"84.6%"},"code_links":[{"title":"erjui/DHO","url":"https://github.com/erjui/DHO"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/knowledge-distillation-on-imagenet","task":"Knowledge Distillation","dataset_variant":"ImageNet","rows":52,"metrics":["Top-1 accuracy %","model size","CRD training setting"],"first_row_in_archive_order":{"model":"ScaleKD (T:BEiT-L S:ViT-B/14)","paper":"/paper/scalekd-strong-vision-transformers-could-be","metrics":{"CRD training setting":"✘","Top-1 accuracy %":"86.43","model size":"87M"},"code_links":[{"title":"deep-optimization/scalekd","url":"https://github.com/deep-optimization/scalekd"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-classification-on-imagenet-v2","task":"Image Classification","dataset_variant":"ImageNet V2","rows":33,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"Model soups (BASIC-L)","paper":"/paper/model-soups-averaging-weights-of-multiple","metrics":{"Top 1 Accuracy":"84.63"},"code_links":[{"title":"mlfoundations/model-soups","url":"https://github.com/mlfoundations/model-soups"},{"title":"Burf/ModelSoups","url":"https://github.com/Burf/ModelSoups"},{"title":"facebookresearch/ModelRatatouille","url":"https://github.com/facebookresearch/ModelRatatouille"},{"title":"hwk0702/keras2torch","url":"https://github.com/hwk0702/keras2torch/tree/main/Computer_Vision/Model_Soup"},{"title":"flowritecom/flow-merge","url":"https://github.com/flowritecom/flow-merge"},{"title":"shallowlearn/sportsreid","url":"https://github.com/shallowlearn/sportsreid"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/quantization-on-imagenet","task":"Quantization","dataset_variant":"ImageNet","rows":27,"metrics":["Top-1 Accuracy (%)","Weight bits","Activation bits"],"first_row_in_archive_order":{"model":"FQ-ViT (ViT-L)","paper":"/paper/fq-vit-fully-quantized-vision-transformer","metrics":{"Activation bits":"8","Top-1 Accuracy (%)":"85.03","Weight bits":"8"},"code_links":[{"title":"megvii-research/FQ-ViT","url":"https://github.com/megvii-research/FQ-ViT"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/zero-shot-transfer-image-classification-on-1","task":"Zero-Shot Transfer Image Classification","dataset_variant":"ImageNet","rows":23,"metrics":["Param","Accuracy (Private)","Accuracy (Public)"],"first_row_in_archive_order":{"model":"M2-Encoder","paper":"/paper/boldsymbol-m-2-encoder-advancing-bilingual","metrics":{"Accuracy (Private)":"88.5","Param":"10B"},"code_links":[{"title":"alipay/Ant-Multi-Modal-Framework","url":"https://github.com/alipay/Ant-Multi-Modal-Framework/tree/main/prj/M2_Encoder"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/data-augmentation-on-imagenet","task":"Data Augmentation","dataset_variant":"ImageNet","rows":17,"metrics":["Accuracy (%)"],"first_row_in_archive_order":{"model":"DeiT-B (+MixPro)","paper":"/paper/mixpro-data-augmentation-with-maskmix-and","metrics":{"Accuracy (%)":"82.9"},"code_links":[{"title":"fistyee/mixpro","url":"https://github.com/fistyee/mixpro"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/network-pruning-on-imagenet","task":"Network Pruning","dataset_variant":"ImageNet","rows":16,"metrics":["Accuracy","GFLOPs","MParams"],"first_row_in_archive_order":{"model":"ResNet50-2.3 GFLOPs","paper":"/paper/pruning-filters-for-efficient-convnets","metrics":{"Accuracy":"78.79","GFLOPs":"2.335","MParams":"14.811"},"code_links":[{"title":"PaddlePaddle/PaddleOCR","url":"https://github.com/PaddlePaddle/PaddleOCR"},{"title":"VainF/Torch-Pruning","url":"https://github.com/VainF/Torch-Pruning"},{"title":"he-y/filter-pruning-geometric-median","url":"https://github.com/he-y/filter-pruning-geometric-median"},{"title":"midasklr/yolov5prune","url":"https://github.com/midasklr/yolov5prune"},{"title":"oandrienko/fast-semantic-segmentation","url":"https://github.com/oandrienko/fast-semantic-segmentation"},{"title":"marcoancona/TorchPruner","url":"https://github.com/marcoancona/TorchPruner"},{"title":"mingsun-tse/regularization-pruning","url":"https://github.com/mingsun-tse/regularization-pruning"},{"title":"Adlik/model_optimizer","url":"https://github.com/Adlik/model_optimizer"},{"title":"guoxiaolu/model_compression","url":"https://github.com/guoxiaolu/model_compression"},{"title":"matthew-mcateer/Keras_pruning","url":"https://github.com/matthew-mcateer/Keras_pruning"},{"title":"AlumLuther/PruningFilters","url":"https://github.com/AlumLuther/PruningFilters"},{"title":"EstherBear/implementation-of-pruning-filters","url":"https://github.com/EstherBear/implementation-of-pruning-filters"},{"title":"arturjordao/PruningNeuralNetworks","url":"https://github.com/arturjordao/PruningNeuralNetworks"},{"title":"lehduong/ginp","url":"https://github.com/lehduong/ginp"},{"title":"lehduong/kesi","url":"https://github.com/lehduong/kesi"},{"title":"siyuan0/pytorch_model_prune","url":"https://github.com/siyuan0/pytorch_model_prune"},{"title":"mvpzhangqiu/yolov5prune","url":"https://github.com/mvpzhangqiu/yolov5prune"},{"title":"cailinhang/2018-Graduation-Project","url":"https://github.com/cailinhang/2018-Graduation-Project"},{"title":"AnishDelft/ModelCompression","url":"https://github.com/AnishDelft/ModelCompression"},{"title":"mattangus/fast-semantic-segmentation","url":"https://github.com/mattangus/fast-semantic-segmentation"},{"title":"prerakmody/CS4180-DL","url":"https://github.com/prerakmody/CS4180-DL"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-reconstruction-on-imagenet","task":"Image Reconstruction","dataset_variant":"ImageNet","rows":15,"metrics":["FID","LPIPS","PSNR","SSIM"],"first_row_in_archive_order":{"model":"MGVQ (16x16x8)","paper":"/paper/mgvq-could-vq-vae-beat-vae-a-generalizable","metrics":{"FID":"0.49","LPIPS":"0.086","PSNR":"24.70","SSIM":"0.787"},"code_links":[{"title":"MKJia/MGVQ","url":"https://github.com/MKJia/MGVQ"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/prompt-engineering-on-imagenet","task":"Prompt Engineering","dataset_variant":"ImageNet","rows":15,"metrics":["Harmonic mean"],"first_row_in_archive_order":{"model":"PromptKD","paper":"/paper/promptkd-unsupervised-prompt-distillation-for","metrics":{"Harmonic mean":"77.62"},"code_links":[{"title":"zhengli97/promptkd","url":"https://github.com/zhengli97/promptkd"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/contrastive-learning-on-imagenet-1k","task":"Contrastive Learning","dataset_variant":"imagenet-1k","rows":14,"metrics":["ImageNet Top-1 Accuracy"],"first_row_in_archive_order":{"model":"ResNet50","paper":"/paper/kernel-ssl-kernel-kl-divergence-for-self","metrics":{"ImageNet Top-1 Accuracy":"73.6"},"code_links":[{"title":"yifanzhang-pro/matrix-ssl","url":"https://github.com/yifanzhang-pro/matrix-ssl"},{"title":"yifanzhang-pro/matrix-llm","url":"https://github.com/yifanzhang-pro/matrix-llm"},{"title":"huang-research-group/Matrix-SSL","url":"https://github.com/huang-research-group/Matrix-SSL"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/zero-shot-transfer-image-classification-on-3","task":"Zero-Shot Transfer Image Classification","dataset_variant":"ImageNet V2","rows":13,"metrics":["Accuracy (Private)","Accuracy (Public)"],"first_row_in_archive_order":{"model":"BASIC (Lion)","paper":null,"metrics":{"Accuracy (Private)":"81.2"},"code_links":[]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-clustering-on-imagenet","task":"Image Clustering","dataset_variant":"ImageNet","rows":12,"metrics":["Accuracy","NMI","ARI"],"first_row_in_archive_order":{"model":"TURTLE (CLIP + DINOv2)","paper":"/paper/let-go-of-your-labels-with-unsupervised-1","metrics":{"ARI":"62.5","Accuracy":"72.9","NMI":"88.2"},"code_links":[{"title":"mlbio-epfl/turtle","url":"https://github.com/mlbio-epfl/turtle"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/model-compression-on-imagenet","task":"Model Compression","dataset_variant":"ImageNet","rows":12,"metrics":["Top-1"],"first_row_in_archive_order":{"model":"ADLIK-MO-ResNet50+W4A4","paper":"/paper/learned-step-size-quantization","metrics":{"Top-1":"77.878"},"code_links":[{"title":"zhutmost/lsq-net","url":"https://github.com/zhutmost/lsq-net"},{"title":"hustzxd/LSQuantization","url":"https://github.com/hustzxd/LSQuantization"},{"title":"ZouJiu1/LSQplus","url":"https://github.com/ZouJiu1/LSQplus"},{"title":"Adlik/model_optimizer","url":"https://github.com/Adlik/model_optimizer"},{"title":"DeadAt0m/LSQFakeQuantize-PyTorch","url":"https://github.com/DeadAt0m/LSQFakeQuantize-PyTorch"},{"title":"DeadAt0m/LSQ-PyTorch","url":"https://github.com/DeadAt0m/LSQ-PyTorch"},{"title":"Shunli-Wang/Tiny-YOLO-LSQ","url":"https://github.com/Shunli-Wang/Tiny-YOLO-LSQ"},{"title":"Kelvinyu1117/LSQ-implementation","url":"https://github.com/Kelvinyu1117/LSQ-implementation"},{"title":"jiyoonkm/columnquant","url":"https://github.com/jiyoonkm/columnquant"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/zero-shot-composed-image-retrieval-zs-cir-on-5","task":"Zero-Shot Composed Image Retrieval (ZS-CIR)","dataset_variant":"ImageNet","rows":11,"metrics":["Average Recall"],"first_row_in_archive_order":{"model":"iSEARLE-XL (CLIP L/14)","paper":"/paper/isearle-improving-textual-inversion-for-zero","metrics":{"Average Recall":"24.46"},"code_links":[{"title":"miccunifi/searle","url":"https://github.com/miccunifi/searle"},{"title":"miccunifi/circo","url":"https://github.com/miccunifi/circo"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/sparse-learning-on-imagenet","task":"Sparse Learning","dataset_variant":"ImageNet","rows":9,"metrics":["Top-1 Accuracy"],"first_row_in_archive_order":{"model":"Resnet-50: 80% Sparse","paper":"/paper/rigging-the-lottery-making-all-tickets-1","metrics":{"Top-1 Accuracy":"77.1"},"code_links":[{"title":"google-research/rigl","url":"https://github.com/google-research/rigl"},{"title":"verbose-avocado/rigl-torch","url":"https://github.com/verbose-avocado/rigl-torch"},{"title":"nollied/rigl-torch","url":"https://github.com/nollied/rigl-torch"},{"title":"hyeon95y/sparselinear","url":"https://github.com/hyeon95y/sparselinear"},{"title":"Shiweiliuiiiiiii/In-Time-Over-Parameterization","url":"https://github.com/Shiweiliuiiiiiii/In-Time-Over-Parameterization"},{"title":"vita-group/granet","url":"https://github.com/vita-group/granet"},{"title":"Shiweiliuiiiiiii/GraNet","url":"https://github.com/Shiweiliuiiiiiii/GraNet"},{"title":"varun19299/rigl-reproducibility","url":"https://github.com/varun19299/rigl-reproducibility"},{"title":"calgaryml/condensed-sparsity","url":"https://github.com/calgaryml/condensed-sparsity"},{"title":"stevenboys/moon","url":"https://github.com/stevenboys/moon"},{"title":"stevenboys/agent","url":"https://github.com/stevenboys/agent"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/unsupervised-image-classification-on-imagenet","task":"Unsupervised Image Classification","dataset_variant":"ImageNet","rows":9,"metrics":["Accuracy (%)","ARI"],"first_row_in_archive_order":{"model":"TURTLE (CLIP + DINOv2)","paper":"/paper/let-go-of-your-labels-with-unsupervised-1","metrics":{"ARI":"62.5","Accuracy (%)":"72.9"},"code_links":[{"title":"mlbio-epfl/turtle","url":"https://github.com/mlbio-epfl/turtle"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/feature-upsampling-on-imagenet","task":"Feature Upsampling","dataset_variant":"ImageNet","rows":8,"metrics":["ADCC","Average Drop","Average Increase"],"first_row_in_archive_order":{"model":"JAFAR","paper":"/paper/jafar-jack-up-any-feature-at-any-resolution-1","metrics":{"ADCC":"73.3","Average Drop":"17.4","Average Increase":"30.9"},"code_links":[{"title":"PaulCouairon/JAFAR","url":"https://github.com/PaulCouairon/JAFAR"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/few-shot-image-classification-on-imagenet-1-1","task":"Few-Shot Image Classification","dataset_variant":"ImageNet - 1-shot","rows":8,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"ViT-MoE-15B (Every-2)","paper":"/paper/scaling-vision-with-sparse-mixture-of-experts","metrics":{"Top 1 Accuracy":"68.66"},"code_links":[{"title":"google-research/vmoe","url":"https://github.com/google-research/vmoe"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/few-shot-image-classification-on-imagenet-5","task":"Few-Shot Image Classification","dataset_variant":"ImageNet - 5-shot","rows":8,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"ViT-MoE-15B (Every-2)","paper":"/paper/scaling-vision-with-sparse-mixture-of-experts","metrics":{"Top 1 Accuracy":"82.78"},"code_links":[{"title":"google-research/vmoe","url":"https://github.com/google-research/vmoe"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/prompt-engineering-on-imagenet-v2","task":"Prompt Engineering","dataset_variant":"ImageNet V2","rows":8,"metrics":["Top-1 accuracy %"],"first_row_in_archive_order":{"model":"HPT++","paper":"/paper/hpt-hierarchically-prompting-vision-language","metrics":{"Top-1 accuracy %":"65.31"},"code_links":[{"title":"vill-lab/2024-aaai-hpt","url":"https://github.com/vill-lab/2024-aaai-hpt"},{"title":"ThomasWangY/2024-AAAI-HPT","url":"https://github.com/ThomasWangY/2024-AAAI-HPT"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/few-shot-image-classification-on-imagenet-10","task":"Few-Shot Image Classification","dataset_variant":"ImageNet - 10-shot","rows":7,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"MAWS (ViT-6.5B)","paper":"/paper/the-effectiveness-of-mae-pre-pretraining-for","metrics":{"Top 1 Accuracy":"84.6"},"code_links":[{"title":"facebookresearch/maws","url":"https://github.com/facebookresearch/maws"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-super-resolution-on-imagenet","task":"Image Super-Resolution","dataset_variant":"ImageNet","rows":6,"metrics":["FID","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DAVI","paper":"/paper/diffusion-prior-based-amortized-variational","metrics":{"FID":"36.27","PSNR":"26.58"},"code_links":[{"title":"mlvlab/davi","url":"https://github.com/mlvlab/davi"},{"title":"kdhRick2222/Exposure-slot","url":"https://github.com/kdhRick2222/Exposure-slot"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/jpeg-decompression-on-imagenet","task":"JPEG Decompression","dataset_variant":"ImageNet","rows":6,"metrics":["FID-5K","IS","CA","PD"],"first_row_in_archive_order":{"model":"Palette (QF: 20)","paper":"/paper/palette-image-to-image-diffusion-models-1","metrics":{"CA":"73.5","FID-5K":"4.3","IS":"208.7","PD":"37.1"},"code_links":[{"title":"Janspiry/Palette-Image-to-Image-Diffusion-Models","url":"https://github.com/Janspiry/Palette-Image-to-Image-Diffusion-Models"},{"title":"LouisRouss/Diffusion-Based-Model-for-Colorization","url":"https://github.com/LouisRouss/Diffusion-Based-Model-for-Colorization"},{"title":"crosszamirski/guided-i2i","url":"https://github.com/crosszamirski/guided-i2i"},{"title":"kylelo/roofdiffusion","url":"https://github.com/kylelo/roofdiffusion"},{"title":"omerb01/puq","url":"https://github.com/omerb01/puq"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/weakly-supervised-object-localization-on-2","task":"Weakly-Supervised Object Localization","dataset_variant":"ImageNet","rows":6,"metrics":["GT-known localization accuracy","Top-1 Localization Accuracy","average top-1 classification accuracy"],"first_row_in_archive_order":{"model":"Stable diffusion","paper":"/paper/generative-prompt-model-for-weakly-supervised","metrics":{"GT-known localization accuracy":"75.0","Top-1 Localization Accuracy":"65.2"},"code_links":[{"title":"callsys/genpromp","url":"https://github.com/callsys/genpromp"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/few-shot-image-classification-on-imagenet-0","task":"Few-Shot Image Classification","dataset_variant":"ImageNet - 0-Shot","rows":5,"metrics":["Accuracy"],"first_row_in_archive_order":{"model":"DebiasPL (ResNet50)","paper":"/paper/debiased-learning-from-naturally-imbalanced","metrics":{"Accuracy":"68.3%"},"code_links":[{"title":"frank-xwang/debiased-pseudo-labeling","url":"https://github.com/frank-xwang/debiased-pseudo-labeling"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-inpainting-on-imagenet","task":"Image Inpainting","dataset_variant":"ImageNet","rows":5,"metrics":["FID","PSNR","SSIM"],"first_row_in_archive_order":{"model":"WavePaint","paper":"/paper/wavepaint-resource-efficient-token-mixer-for","metrics":{"FID":"3.21"},"code_links":[{"title":"pranavphoenix/WavePaint","url":"https://github.com/pranavphoenix/WavePaint"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/adversarial-robustness-on-imagenet","task":"Adversarial Robustness","dataset_variant":"ImageNet","rows":4,"metrics":["Accuracy"],"first_row_in_archive_order":{"model":"ResNet-50 (SGD, Cosine)","paper":"/paper/are-transformers-more-robust-than-cnns","metrics":{"Accuracy":"77.4"},"code_links":[{"title":"ytongbai/ViTs-vs-CNNs","url":"https://github.com/ytongbai/ViTs-vs-CNNs"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-classification-with-dp-on-imagenet","task":"Image Classification with Differential Privacy","dataset_variant":"ImageNet","rows":4,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"NFResnet-50","paper":"/paper/tan-without-a-burn-scaling-laws-of-dp-sgd","metrics":{"Top 1 Accuracy":"39.2"},"code_links":[{"title":"facebookresearch/tan","url":"https://github.com/facebookresearch/tan"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-colorization-on-imagenet","task":"Image Colorization","dataset_variant":"ImageNet","rows":4,"metrics":["Consistency","FID"],"first_row_in_archive_order":{"model":"DDRM","paper":"/paper/zero-shot-image-restoration-using-denoising","metrics":{"Consistency":"260.4","FID":"36.56"},"code_links":[{"title":"wyhuai/ddnm","url":"https://github.com/wyhuai/ddnm"},{"title":"xypeng9903/k-diffusion-inverse-problems","url":"https://github.com/xypeng9903/k-diffusion-inverse-problems"},{"title":"ipc-lab/deepjscc-diffusion","url":"https://github.com/ipc-lab/deepjscc-diffusion"},{"title":"andreamazzitelli/ProjectNN","url":"https://github.com/andreamazzitelli/ProjectNN"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/weakly-supervised-object-detection-on","task":"Weakly Supervised Object Detection","dataset_variant":"ImageNet","rows":4,"metrics":["MAP"],"first_row_in_archive_order":{"model":"PCL-OB-G-Ens + FRCNN","paper":"/paper/pcl-proposal-cluster-learning-for-weakly","metrics":{"MAP":"19.6"},"code_links":[{"title":"ppengtang/pcl.pytorch","url":"https://github.com/ppengtang/pcl.pytorch"},{"title":"ppengtang/oicr","url":"https://github.com/ppengtang/oicr"},{"title":"JoegameZhou/mPanGu-Alpha-53","url":"https://github.com/JoegameZhou/mPanGu-Alpha-53"},{"title":"George-Holbrow-Wilshaw/Human-Protein-Atlas-Kaggle","url":"https://github.com/George-Holbrow-Wilshaw/Human-Protein-Atlas-Kaggle"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/adversarial-defense-on-imagenet","task":"Adversarial Defense","dataset_variant":"ImageNet","rows":3,"metrics":["Accuracy"],"first_row_in_archive_order":{"model":"ResNet101","paper":"/paper/nomaro-defending-against-adversarial-attacks","metrics":{"Accuracy":"99.8%"},"code_links":[{"title":"as791/NOMARO_defense","url":"https://github.com/as791/NOMARO_defense"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-deblurring-on-imagenet","task":"Image Deblurring","dataset_variant":"ImageNet","rows":3,"metrics":["FID","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DDNM","paper":"/paper/zero-shot-image-restoration-using-denoising","metrics":{"FID":"1.15","PSNR":"44.93","SSIM":"0.994"},"code_links":[{"title":"wyhuai/ddnm","url":"https://github.com/wyhuai/ddnm"},{"title":"xypeng9903/k-diffusion-inverse-problems","url":"https://github.com/xypeng9903/k-diffusion-inverse-problems"},{"title":"ipc-lab/deepjscc-diffusion","url":"https://github.com/ipc-lab/deepjscc-diffusion"},{"title":"andreamazzitelli/ProjectNN","url":"https://github.com/andreamazzitelli/ProjectNN"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/semi-supervised-image-classification-on-16","task":"Semi-Supervised Image Classification","dataset_variant":"ImageNet - 0.2% labeled data","rows":3,"metrics":["ImageNet Top-1 Accuracy"],"first_row_in_archive_order":{"model":"DebiasPL (ResNet-50)","paper":"/paper/debiased-learning-from-naturally-imbalanced","metrics":{"ImageNet Top-1 Accuracy":"69.6%"},"code_links":[{"title":"frank-xwang/debiased-pseudo-labeling","url":"https://github.com/frank-xwang/debiased-pseudo-labeling"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/color-image-denoising-on-imagenet-sigma100","task":"Color Image Denoising","dataset_variant":"ImageNet sigma100","rows":2,"metrics":["LPIPS","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DMID-p","paper":"/paper/stimulating-the-diffusion-model-for-image","metrics":{"LPIPS":"0.156","PSNR":"24.61","SSIM":"0.7987"},"code_links":[{"title":"li-tong-621/dmid","url":"https://github.com/li-tong-621/dmid"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/color-image-denoising-on-imagenet-sigma150","task":"Color Image Denoising","dataset_variant":"ImageNet sigma150","rows":2,"metrics":["LPIPS","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DMID-p","paper":"/paper/stimulating-the-diffusion-model-for-image","metrics":{"LPIPS":"0.259","PSNR":"22.94","SSIM":"0.6932"},"code_links":[{"title":"li-tong-621/dmid","url":"https://github.com/li-tong-621/dmid"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/color-image-denoising-on-imagenet-sigma200","task":"Color Image Denoising","dataset_variant":"ImageNet sigma200","rows":2,"metrics":["LPIPS","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DMID-p","paper":"/paper/stimulating-the-diffusion-model-for-image","metrics":{"LPIPS":"0.259","PSNR":"21.56","SSIM":"0.6932"},"code_links":[{"title":"li-tong-621/dmid","url":"https://github.com/li-tong-621/dmid"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/color-image-denoising-on-imagenet-sigma250","task":"Color Image Denoising","dataset_variant":"ImageNet sigma250","rows":2,"metrics":["LPIPS","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DMID-p","paper":"/paper/stimulating-the-diffusion-model-for-image","metrics":{"LPIPS":"0.289","PSNR":"20.87","SSIM":"0.6701"},"code_links":[{"title":"li-tong-621/dmid","url":"https://github.com/li-tong-621/dmid"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/color-image-denoising-on-imagenet-sigma50","task":"Color Image Denoising","dataset_variant":"ImageNet sigma50","rows":2,"metrics":["LPIPS","PSNR","SSIM"],"first_row_in_archive_order":{"model":"DMID-p","paper":"/paper/stimulating-the-diffusion-model-for-image","metrics":{"LPIPS":"0.087","PSNR":"27.59","SSIM":"0.8722"},"code_links":[{"title":"li-tong-621/dmid","url":"https://github.com/li-tong-621/dmid"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/medical-image-classification-on-imagenet","task":"Medical Image Classification","dataset_variant":"ImageNet","rows":2,"metrics":["GFLOPs","Top 1 Accuracy"],"first_row_in_archive_order":{"model":"DaViT-T","paper":"/paper/davit-dual-attention-vision-transformers","metrics":{"GFLOPs":"4.5"},"code_links":[{"title":"rwightman/pytorch-image-models","url":"https://github.com/rwightman/pytorch-image-models"},{"title":"leondgarse/keras_cv_attention_models","url":"https://github.com/leondgarse/keras_cv_attention_models/tree/main/keras_cv_attention_models/davit"},{"title":"dingmyu/davit","url":"https://github.com/dingmyu/davit"},{"title":"birder/birder","url":"https://gitlab.com/birder/birder"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/zero-shot-learning-on-imagenet","task":"Zero-Shot Learning","dataset_variant":"ImageNet","rows":2,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"ZLaP","paper":"/paper/label-propagation-for-zero-shot","metrics":{"Top 1 Accuracy":"72.1"},"code_links":[{"title":"vladan-stojnic/zlap","url":"https://github.com/vladan-stojnic/zlap"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-classification-on-imagenet-1k","task":"Image Classification","dataset_variant":"imagenet-1k","rows":1,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"BinaryViT","paper":"/paper/binaryvit-pushing-binary-vision-transformers","metrics":{"Top 1 Accuracy":"70.6"},"code_links":[{"title":"phuoc-hoan-le/binaryvit","url":"https://github.com/phuoc-hoan-le/binaryvit"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-clustering-on-imagenet-1k","task":"Image Clustering","dataset_variant":"imagenet-1k","rows":1,"metrics":["ARI","Accuracy","NMI"],"first_row_in_archive_order":{"model":"TAC","paper":"/paper/image-clustering-with-external-guidance","metrics":{"ARI":"0.435","Accuracy":"0.582","NMI":"0.799"},"code_links":[{"title":"xlearning-scu/2024-icml-tac","url":"https://github.com/xlearning-scu/2024-icml-tac"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/image-segmentation-on-imagenet","task":"Image Segmentation","dataset_variant":"ImageNet","rows":1,"metrics":["GFLOPs"],"first_row_in_archive_order":{"model":"MobileOne-S0","paper":"/paper/an-improved-one-millisecond-mobile-backbone","metrics":{"GFLOPs":"0.275"},"code_links":[{"title":"rwightman/pytorch-image-models","url":"https://github.com/rwightman/pytorch-image-models"},{"title":"PaddlePaddle/PaddleDetection","url":"https://github.com/PaddlePaddle/PaddleDetection"},{"title":"BR-IDL/PaddleViT","url":"https://github.com/BR-IDL/PaddleViT/tree/develop/image_classification/MobileOne"},{"title":"apple/ml-mobileone","url":"https://github.com/apple/ml-mobileone"},{"title":"frgfm/Holocron","url":"https://github.com/frgfm/Holocron"},{"title":"chengpengchen/repghost","url":"https://github.com/chengpengchen/repghost"},{"title":"yakhyo/gaze-estimation","url":"https://github.com/yakhyo/gaze-estimation"},{"title":"federicopozzi33/MobileOne-PyTorch","url":"https://github.com/federicopozzi33/MobileOne-PyTorch"},{"title":"james77777778/keras-image-models","url":"https://github.com/james77777778/keras-image-models"},{"title":"birder/birder","url":"https://gitlab.com/birder/birder"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/transductive-zero-shot-classification-on","task":"Transductive Zero-Shot Classification","dataset_variant":"ImageNet","rows":1,"metrics":["Top 1 Accuracy"],"first_row_in_archive_order":{"model":"ZLaP","paper":"/paper/label-propagation-for-zero-shot","metrics":{"Top 1 Accuracy":"72.7"},"code_links":[{"title":"vladan-stojnic/zlap","url":"https://github.com/vladan-stojnic/zlap"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/visual-question-answering-vqa-on-imagenet","task":"Visual Question Answering (VQA)","dataset_variant":"ImageNet","rows":1,"metrics":["ClipMatch@1","ClipMatch@5","Contains","ExactMatch","Follow-up ClipMatch@1","Follow-up ClipMatch@5","Follow-up Contains","Follow-up ExactMatch"],"first_row_in_archive_order":{"model":"BLIP-2 OPT","paper":"/paper/open-ended-vqa-benchmarking-of-vision","metrics":{"ClipMatch@1":"57.10","ClipMatch@5":"77.24","Contains":"35.49","ExactMatch":"0.87","Follow-up ClipMatch@1":"67.22","Follow-up ClipMatch@5":"83.54","Follow-up Contains":"40.31","Follow-up ExactMatch":"2.54"},"code_links":[{"title":"lmb-freiburg/ovqa","url":"https://github.com/lmb-freiburg/ovqa"}]},"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/classification-on-imagenet-1k","task":"Classification","dataset_variant":"imagenet-1k","rows":0,"metrics":["Accuracy"],"first_row_in_archive_order":null,"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"},{"leaderboard":"/sota/semi-supervised-image-classification-on-27","task":"Semi-Supervised Image Classification","dataset_variant":"ImageNet","rows":0,"metrics":["Accuracy at 1%"],"first_row_in_archive_order":null,"note":"rows are the archive's own order at snapshot; nothing here re-ranks them"}],"papers_with_a_benchmark_row":[{"paper":"/paper/mgvq-could-vq-vae-beat-vae-a-generalizable","title":"MGVQ: Could VQ-VAE Beat VAE? A Generalizable Tokenizer with Multi-group Quantization","date":"2025-07-14","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/jafar-jack-up-any-feature-at-any-resolution-1","title":"JAFAR: Jack up Any Feature at Any Resolution","date":"2025-06-10","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/sst-self-training-with-self-adaptive-1","title":"SST: Self-training with Self-adaptive Thresholding for Semi-supervised Learning","date":"2025-05-31","rows_on_this_dataset":10,"code_links":0,"syntology":null},{"paper":"/paper/mmrl-parameter-efficient-and-interaction","title":"MMRL++: Parameter-Efficient and Interaction-Aware Representation Learning for Vision-Language Models","date":"2025-05-15","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/simple-semi-supervised-knowledge-distillation","title":"Simple Semi-supervised Knowledge Distillation from Vision-Language Models via $\\mathbf{\\texttt{D}}$ual-$\\mathbf{\\texttt{H}}$ead $\\mathbf{\\texttt{O}}$ptimization","date":"2025-05-12","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/gigatok-scaling-visual-tokenizers-to-3","title":"GigaTok: Scaling Visual Tokenizers to 3 Billion Parameters for Autoregressive Image Generation","date":"2025-04-11","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":10,"samples_unverified":5,"pointer_only_for_licence":15,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/enhanced-ood-detection-through-cross-modal","title":"Enhanced OoD Detection through Cross-Modal Alignment of Multi-Modal Representations","date":"2025-03-24","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/exploring-a-principled-framework-for-deep-1","title":"Exploring a Principled Framework for Deep Subspace Clustering","date":"2025-03-21","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":4,"samples_unverified":0,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mmrl-multi-modal-representation-learning-for","title":"MMRL: Multi-Modal Representation Learning for Vision-Language Models","date":"2025-03-11","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/compress-image-to-patches-for-vision","title":"Compress image to patches for Vision Transformer","date":"2025-02-14","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/learnable-polynomial-trigonometric-and","title":"Polynomial, trigonometric, and tropical activations","date":"2025-02-03","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/spectralkd-understanding-and-optimizing","title":"SpectralKD: A Unified Framework for Interpreting and Distilling Vision Transformers via Spectral Analysis","date":"2024-12-26","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/mamba2d-a-natively-multi-dimensional-state","title":"Mamba2D: A Natively Multi-Dimensional State-Space Model for Vision Tasks","date":"2024-12-20","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/preventing-local-pitfalls-in-vector","title":"Preventing Local Pitfalls in Vector Quantization via Optimal Transport","date":"2024-12-19","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/taming-scalable-visual-tokenizer-for","title":"Taming Scalable Visual Tokenizer for Autoregressive Image Generation","date":"2024-12-03","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":13,"samples_ran":9,"samples_unverified":4,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/on-the-performance-analysis-of-momentum","title":"On the Performance Analysis of Momentum Method: A Frequency Domain Perspective","date":"2024-11-29","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/multimodal-autoregressive-pre-training-of","title":"Multimodal Autoregressive Pre-training of Large Vision Encoders","date":"2024-11-21","rows_on_this_dataset":6,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/scalekd-strong-vision-transformers-could-be","title":"ScaleKD: Strong Vision Transformers Could Be Excellent Teachers","date":"2024-11-11","rows_on_this_dataset":5,"code_links":1,"syntology":null},{"paper":"/paper/performance-of-gaussian-mixture-model","title":"Performance of Gaussian Mixture Model Classifiers on Embedded Feature Spaces","date":"2024-10-17","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/stabilize-the-latent-space-for-image","title":"Stabilize the Latent Space for Image Autoregressive Modeling: A Unified Perspective","date":"2024-10-16","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":3,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/debiformer-vision-transformer-with-deformable","title":"DeBiFormer: Vision Transformer with Deformable Agent Bi-level Routing Attention","date":"2024-10-11","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/synco-synthetic-hard-negatives-in-contrastive","title":"SynCo: Synthetic Hard Negatives in Contrastive Learning for Better Unsupervised Visual Representations","date":"2024-10-03","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/hvt-a-comprehensive-vision-framework-for","title":"HVT: A Comprehensive Vision Framework for Learning in Non-Euclidean Space","date":"2024-09-25","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/maskbit-embedding-free-image-generation-via","title":"MaskBit: Embedding-free Image Generation via Bit Tokens","date":"2024-09-24","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":8,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/kolmogorov-arnold-transformer","title":"Kolmogorov-Arnold Transformer","date":"2024-09-16","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/open-magvit2-an-open-source-project-toward","title":"Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation","date":"2024-09-06","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/hpt-hierarchically-prompting-vision-language","title":"HPT++: Hierarchically Prompting Vision-Language Models with Multi-Granularity Knowledge Generation and Improved Structure Modeling","date":"2024-08-27","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/cas-vit-convolutional-additive-self-attention","title":"CAS-ViT: Convolutional Additive Self-attention Vision Transformers for Efficient Mobile Applications","date":"2024-08-07","rows_on_this_dataset":4,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/2408-02014","title":"Unsupervised Representation Learning by Balanced Self Attention Matching","date":"2024-08-04","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/diffusion-prior-based-amortized-variational","title":"Diffusion Prior-Based Amortized Variational Inference for Noisy Inverse Problems","date":"2024-07-23","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":27,"samples_ran":21,"samples_unverified":6,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/colormae-exploring-data-independent-masking","title":"ColorMAE: Exploring data-independent masking strategies in Masked AutoEncoders","date":"2024-07-17","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/unconstrained-open-vocabulary-image","title":"Unconstrained Open Vocabulary Image Classification: Zero-Shot Transfer from Text to Image via CLIP Inversion","date":"2024-07-15","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/mambavision-a-hybrid-mamba-transformer-vision","title":"MambaVision: A Hybrid Mamba-Transformer Vision Backbone","date":"2024-07-10","rows_on_this_dataset":7,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":2,"samples_unverified":6,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-refreshed-similarity-based-upsampler-for","title":"A Refreshed Similarity-based Upsampler for Direct High-Ratio Feature Upsampling","date":"2024-07-02","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":6,"samples_unverified":2,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/scaling-the-codebook-size-of-vqgan-to-100000","title":"Scaling the Codebook Size of VQGAN to 100,000 with a Utilization Rate of 99%","date":"2024-06-17","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/let-go-of-your-labels-with-unsupervised-1","title":"Let Go of Your Labels with Unsupervised Transfer","date":"2024-06-11","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/an-image-is-worth-32-tokens-for","title":"An Image is Worth 32 Tokens for Reconstruction and Generation","date":"2024-06-11","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/dino-as-a-von-mises-fisher-mixture-model-1","title":"DINO as a von Mises-Fisher mixture model","date":"2024-05-17","rows_on_this_dataset":3,"code_links":0,"syntology":null},{"paper":"/paper/isearle-improving-textual-inversion-for-zero","title":"iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval","date":"2024-05-05","rows_on_this_dataset":4,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/ghostnetv3-exploring-the-training-strategies","title":"GhostNetV3: Exploring the Training Strategies for Compact Models","date":"2024-04-17","rows_on_this_dataset":4,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mobilenetv4-universal-models-for-the-mobile","title":"MobileNetV4 -- Universal Models for the Mobile Ecosystem","date":"2024-04-16","rows_on_this_dataset":5,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":13,"samples_ran":6,"samples_unverified":7,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/label-propagation-for-zero-shot","title":"Label Propagation for Zero-shot Classification with Vision-Language Models","date":"2024-04-05","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/prompt-learning-via-meta-regularization","title":"Prompt Learning via Meta-Regularization","date":"2024-04-01","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":4,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/densenets-reloaded-paradigm-shift-beyond","title":"DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs","date":"2024-03-28","rows_on_this_dataset":5,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tiny-models-are-the-computational-saver-for","title":"Tiny Models are the Computational Saver for Large Models","date":"2024-03-26","rows_on_this_dataset":5,"code_links":1,"syntology":null},{"paper":"/paper/lift-a-surprisingly-simple-lightweight","title":"LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors","date":"2024-03-21","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/featup-a-model-agnostic-framework-for","title":"FeatUp: A Model-Agnostic Framework for Features at Any Resolution","date":"2024-03-15","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/v-kd-improving-knowledge-distillation-using","title":"$V_kD:$ Improving Knowledge Distillation using Orthogonal Projections","date":"2024-03-10","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/promptkd-unsupervised-prompt-distillation-for","title":"PromptKD: Unsupervised Prompt Distillation for Vision-Language Models","date":"2024-03-05","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/hyenapixel-global-image-context-with","title":"HyenaPixel: Global Image Context with Convolutions","date":"2024-02-29","rows_on_this_dataset":5,"code_links":1,"syntology":null},{"paper":"/paper/revit-enhancing-vision-transformers-with","title":"ReViT: Enhancing Vision Transformers Feature Diversity with Attention Residual Connections","date":"2024-02-17","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":24,"samples_ran":22,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/knowledge-distillation-based-on-transformed","title":"Knowledge Distillation Based on Transformed Teacher Matching","date":"2024-02-17","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":5,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mim-refiner-a-contrastive-learning-boost-from","title":"MIM-Refiner: A Contrastive Learning Boost from Intermediate Pre-Trained Representations","date":"2024-02-15","rows_on_this_dataset":7,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/distilled-gradual-pruning-with-pruned-fine","title":"Distilled Gradual Pruning with Pruned Fine-tuning","date":"2024-02-15","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/open-ended-vqa-benchmarking-of-vision","title":"Open-ended VQA benchmarking of Vision-Language models by exploiting Classification datasets and their semantic hierarchy","date":"2024-02-11","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/eva-clip-18b-scaling-clip-to-18-billion","title":"EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters","date":"2024-02-06","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":1,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/boldsymbol-m-2-encoder-advancing-bilingual","title":"M2-Encoder: Advancing Bilingual Image-Text Understanding by Large-scale Efficient Pretraining","date":"2024-01-29","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/scalable-pre-training-of-large-autoregressive","title":"Scalable Pre-training of Large Autoregressive Image Models","date":"2024-01-16","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/omnivec2-a-novel-transformer-based-network","title":"OmniVec2 - A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning","date":"2024-01-01","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/internvl-scaling-up-vision-foundation-models","title":"InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks","date":"2023-12-21","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/exploring-foveation-and-saccade-for-improved","title":"Exploring Foveation and Saccade for Improved Weakly-Supervised Localization","date":"2023-12-16","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/learning-hierarchical-prompt-with-structured","title":"Learning Hierarchical Prompt with Structured Linguistic Knowledge for Vision-Language Models","date":"2023-12-11","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":3,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/graph-convolutions-enrich-the-self-attention","title":"Graph Convolutions Enrich the Self-Attention in Transformers!","date":"2023-12-07","rows_on_this_dataset":4,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":29,"samples_ran":19,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/language-only-efficient-training-of-zero-shot","title":"Language-only Efficient Training of Zero-shot Composed Image Retrieval","date":"2023-12-04","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":3,"samples_unverified":5,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/perceptual-group-tokenizer-building","title":"Perceptual Group Tokenizer: Building Perception with Iterative Grouping","date":"2023-11-30","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/meta-co-training-two-views-are-better-than","title":"Meta Co-Training: Two Views are Better than One","date":"2023-11-29","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/transnext-robust-foveal-visual-perception-for","title":"TransNeXt: Robust Foveal Visual Perception for Vision Transformers","date":"2023-11-28","rows_on_this_dataset":5,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":5,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/beyond-sole-strength-customized-ensembles-for","title":"Beyond Sole Strength: Customized Ensembles for Generalized Vision-Language Models","date":"2023-11-28","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unireplknet-a-universal-perception-large","title":"UniRepLKNet: A Universal Perception Large-Kernel ConvNet for Audio, Video, Point Cloud, Time-Series and Image Recognition","date":"2023-11-27","rows_on_this_dataset":10,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/stable-cluster-discrimination-for-deep-1","title":"Stable Cluster Discrimination for Deep Clustering","date":"2023-11-24","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/gtp-vit-efficient-vision-transformers-via","title":"GTP-ViT: Efficient Vision Transformers via Graph-based Token Propagation","date":"2023-11-06","rows_on_this_dataset":7,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/asymmetric-masked-distillation-for-pre","title":"Asymmetric Masked Distillation for Pre-Training Small Foundation Models","date":"2023-11-06","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/distilling-out-of-distribution-robustness-1","title":"Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models","date":"2023-11-02","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/one-for-all-bridge-the-gap-between-1","title":"One-for-All: Bridge the Gap Between Heterogeneous Architectures in Knowledge Distillation","date":"2023-10-30","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/rckd-response-based-cross-task-knowledge","title":"RCKD: Response-Based Cross-Task Knowledge Distillation for Pathological Image Analysis","date":"2023-10-29","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/sequencematch-revisiting-the-design-of-weak","title":"SequenceMatch: Revisiting the design of weak-strong augmentations for Semi-supervised learning","date":"2023-10-24","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/debiasing-calibrating-and-improving-semi","title":"Debiasing, calibrating, and improving Semi-supervised Learning performance via simple Ensemble Projector","date":"2023-10-24","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/image-clustering-with-external-guidance","title":"Image Clustering with External Guidance","date":"2023-10-18","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":4,"samples_unverified":5,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/semireward-a-general-reward-model-for-semi","title":"SemiReward: A General Reward Model for Semi-supervised Learning","date":"2023-10-04","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vision-transformers-need-registers","title":"Vision Transformers Need Registers","date":"2023-09-28","rows_on_this_dataset":1,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":20,"samples_ran":4,"samples_unverified":16,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/context-i2w-mapping-images-to-context","title":"Context-I2W: Mapping Images to Context-dependent Words for Accurate Zero-Shot Composed Image Retrieval","date":"2023-09-28","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/masked-image-residual-learning-for-scaling-1","title":"Masked Image Residual Learning for Scaling Deeper Vision Transformers","date":"2023-09-25","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/dept-decoupled-prompt-tuning","title":"DePT: Decoupled Prompt Tuning","date":"2023-09-14","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":4,"samples_unverified":4,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/dat-spatially-dynamic-vision-transformer-with","title":"DAT++: Spatially Dynamic Vision Transformer with Deformable Attention","date":"2023-09-04","rows_on_this_dataset":4,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":2,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/read-only-prompt-optimization-for-vision","title":"Read-only Prompt Optimization for Vision-Language Few-shot Learning","date":"2023-08-29","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learning-to-upsample-by-learning-to-sample","title":"Learning to Upsample by Learning to Sample","date":"2023-08-29","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/which-transformer-to-favor-a-comparative","title":"Which Transformer to Favor: A Comparative Analysis of Efficiency in Vision Transformers","date":"2023-08-18","rows_on_this_dataset":15,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":4,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/simmatchv2-semi-supervised-learning-with","title":"SimMatchV2: Semi-Supervised Learning with Graph Consistency","date":"2023-08-13","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/gated-attention-coding-for-training-high","title":"Gated Attention Coding for Training High-performance and Efficient Spiking Neural Networks","date":"2023-08-12","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":7,"samples_unverified":5,"pointer_only_for_licence":12,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/generative-prompt-model-for-weakly-supervised","title":"Generative Prompt Model for Weakly Supervised Object Localization","date":"2023-07-19","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-regulating-prompts-foundational-model","title":"Self-regulating Prompts: Foundational Model Adaptation without Forgetting","date":"2023-07-13","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":21,"samples_ran":7,"samples_unverified":14,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/stimulating-the-diffusion-model-for-image","title":"Stimulating Diffusion Model for Image Denoising via Adaptive Embedding and Ensembling","date":"2023-07-08","rows_on_this_dataset":10,"code_links":1,"syntology":null},{"paper":"/paper/wavepaint-resource-efficient-token-mixer-for","title":"WavePaint: Resource-efficient Token-mixer for Self-supervised Inpainting","date":"2023-07-01","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/binaryvit-pushing-binary-vision-transformers","title":"BinaryViT: Pushing Binary Vision Transformers Towards Convolutional Models","date":"2023-06-29","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/augmenting-sub-model-to-improve-main-model","title":"Masking meets Supervision: A Strong Learning Alliance","date":"2023-06-20","rows_on_this_dataset":6,"code_links":1,"syntology":null},{"paper":"/paper/fastervit-fast-vision-transformers-with","title":"FasterViT: Fast Vision Transformers with Hierarchical Attention","date":"2023-06-09","rows_on_this_dataset":7,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":4,"samples_unverified":5,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/the-information-pathways-hypothesis","title":"The Information Pathways Hypothesis: Transformers are Dynamic Self-Ensembles","date":"2023-06-02","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/hiera-a-hierarchical-vision-transformer","title":"Hiera: A Hierarchical Vision Transformer without the Bells-and-Whistles","date":"2023-06-01","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":0,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/consistency-guided-prompt-learning-for-vision","title":"Consistency-guided Prompt Learning for Vision-Language Models","date":"2023-06-01","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":3,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/kernel-ssl-kernel-kl-divergence-for-self","title":"Matrix Information Theory for Self-Supervised Learning","date":"2023-05-27","rows_on_this_dataset":1,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":5,"samples_unverified":0,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/improving-knowledge-distillation-via-1","title":"Improving Knowledge Distillation via Regularizing Feature Norm and Direction","date":"2023-05-26","rows_on_this_dataset":11,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/knowledge-diffusion-for-distillation-1","title":"Knowledge Diffusion for Distillation","date":"2023-05-25","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/one-peace-exploring-one-general","title":"ONE-PEACE: Exploring One General Representation Model Toward Unlimited Modalities","date":"2023-05-18","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":2,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/alternating-gradient-descent-and-mixture-of","title":"Alternating Gradient Descent and Mixture-of-Experts for Integrated Multimodal Perception","date":"2023-05-10","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/understanding-gaussian-attention-bias-of","title":"Understanding Gaussian Attention Bias of Vision Transformers Using Effective Receptive Fields","date":"2023-05-08","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/mixpro-data-augmentation-with-maskmix-and","title":"MixPro: Data Augmentation with MaskMix and Progressive Attention Labeling for Vision Transformer","date":"2023-04-24","rows_on_this_dataset":12,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":18,"samples_ran":10,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/layernas-neural-architecture-search-in","title":"LayerNAS: Neural Architecture Search in Polynomial Complexity","date":"2023-04-23","rows_on_this_dataset":4,"code_links":0,"syntology":null},{"paper":"/paper/contrastive-tuning-a-little-help-to-make","title":"Contrastive Tuning: A Little Help to Make Masked Autoencoders Forget","date":"2023-04-20","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/dinov2-learning-robust-visual-features","title":"DINOv2: Learning Robust Visual Features without Supervision","date":"2023-04-14","rows_on_this_dataset":7,"code_links":26,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":46,"samples_ran":21,"samples_unverified":25,"pointer_only_for_licence":12,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unicom-universal-and-compact-representation","title":"Unicom: Universal and Compact Representation Learning for Image Retrieval","date":"2023-04-12","rows_on_this_dataset":3,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":3,"samples_unverified":3,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vne-an-effective-method-for-improving-deep","title":"VNE: An Effective Method for Improving Deep Representation by Manipulating Eigenvalue Distribution","date":"2023-04-04","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/rethinking-local-perception-in-lightweight","title":"Rethinking Local Perception in Lightweight Vision Transformer","date":"2023-03-31","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/exploring-the-limits-of-deep-image-clustering","title":"Exploring the Limits of Deep Image Clustering using Pretrained Models","date":"2023-03-31","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":4,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/your-diffusion-model-is-secretly-a-zero-shot","title":"Your Diffusion Model is Secretly a Zero-Shot Classifier","date":"2023-03-28","rows_on_this_dataset":2,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/zero-shot-composed-image-retrieval-with","title":"Zero-Shot Composed Image Retrieval with Textual Inversion","date":"2023-03-27","rows_on_this_dataset":4,"code_links":2,"syntology":null},{"paper":"/paper/eva-clip-improved-training-techniques-for","title":"EVA-CLIP: Improved Training Techniques for CLIP at Scale","date":"2023-03-27","rows_on_this_dataset":2,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":0,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fastvit-a-fast-hybrid-vision-transformer","title":"FastViT: A Fast Hybrid Vision Transformer using Structural Reparameterization","date":"2023-03-24","rows_on_this_dataset":7,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":1,"samples_unverified":4,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/the-effectiveness-of-mae-pre-pretraining-for","title":"The effectiveness of MAE pre-pretraining for billion-scale pretraining","date":"2023-03-23","rows_on_this_dataset":18,"code_links":1,"syntology":null},{"paper":"/paper/visual-representation-learning-from-unlabeled","title":"ViC-MAE: Self-Supervised Representation Learning from Images and Video with Contrastive Masked Autoencoders","date":"2023-03-21","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/mv-mr-multi-views-and-multi-representations","title":"MV-MR: multi-views and multi-representations for self-supervised learning and knowledge distillation","date":"2023-03-21","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/a-closer-look-at-the-training-dynamics-of","title":"Understanding the Role of the Projector in Knowledge Distillation","date":"2023-03-20","rows_on_this_dataset":3,"code_links":4,"syntology":null},{"paper":"/paper/biformer-vision-transformer-with-bi-level","title":"BiFormer: Vision Transformer with Bi-Level Routing Attention","date":"2023-03-15","rows_on_this_dataset":3,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":6,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/r-2-range-regularization-for-model","title":"R2 Loss: Range Restriction Loss for Model Compression and Quantization","date":"2023-03-14","rows_on_this_dataset":13,"code_links":0,"syntology":null},{"paper":"/paper/deepmad-mathematical-architecture-design-for","title":"DeepMAD: Mathematical Architecture Design for Deep Convolutional Neural Network","date":"2023-03-05","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/how-to-use-dropout-correctly-on-residual","title":"How to Use Dropout Correctly on Residual Networks with Batch Normalization","date":"2023-02-13","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/scaling-vision-transformers-to-22-billion","title":"Scaling Vision Transformers to 22 Billion Parameters","date":"2023-02-10","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/pic2word-mapping-pictures-to-words-for-zero","title":"Pic2Word: Mapping Pictures to Words for Zero-shot Composed Image Retrieval","date":"2023-02-06","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learning-customized-visual-models-with","title":"Learning Customized Visual Models with Retrieval-Augmented Knowledge","date":"2023-01-17","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":17,"samples_ran":6,"samples_unverified":11,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/designing-bert-for-convolutional-networks","title":"Designing BERT for Convolutional Networks: Sparse and Hierarchical Masked Modeling","date":"2023-01-09","rows_on_this_dataset":9,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":5,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learning-by-sorting-self-supervised-learning","title":"Learning by Sorting: Self-supervised Learning with Group Ordering Constraints","date":"2023-01-05","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/improving-visual-representation-learning","title":"Improving Visual Representation Learning through Perceptual Understanding","date":"2022-12-30","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/reversible-column-networks","title":"Reversible Column Networks","date":"2022-12-22","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":8,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/from-xception-to-nexception-new-design","title":"From Xception to NEXcepTion: New Design Decisions and Neural Architecture Search","date":"2022-12-16","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/reproducible-scaling-laws-for-contrastive","title":"Reproducible scaling laws for contrastive language-image learning","date":"2022-12-14","rows_on_this_dataset":1,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/efficient-self-supervised-learning-with","title":"Efficient Self-supervised Learning with Contextualized Target Representations for Vision, Speech and Language","date":"2022-12-14","rows_on_this_dataset":1,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":5,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/expeditious-saliency-guided-mix-up-through","title":"Expeditious Saliency-guided Mix-up through Random Gradient Thresholding","date":"2022-12-09","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/co-training-2-l-submodels-for-visual","title":"Co-training $2^L$ Submodels for Visual Recognition","date":"2022-12-09","rows_on_this_dataset":10,"code_links":1,"syntology":null},{"paper":"/paper/learning-domain-invariant-prompt-for-vision","title":"Learning Domain Invariant Prompt for Vision-Language Models","date":"2022-12-08","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/incepformer-efficient-inception-transformer","title":"IncepFormer: Efficient Inception Transformer with Pyramid Pooling for Semantic Segmentation","date":"2022-12-06","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/zero-shot-image-restoration-using-denoising","title":"Zero-Shot Image Restoration Using Denoising Diffusion Null-Space Model","date":"2022-12-01","rows_on_this_dataset":16,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":23,"samples_ran":12,"samples_unverified":11,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pattern-attention-transformer-with-doughnut","title":"Pattern Attention Transformer with Doughnut Kernel","date":"2022-11-30","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/semantic-aware-local-global-vision","title":"Semantic-Aware Local-Global Vision Transformer","date":"2022-11-27","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/differentially-private-image-classification","title":"Differentially Private Image Classification from Features","date":"2022-11-24","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/an-algorithm-for-routing-vectors-in-sequences","title":"An Algorithm for Routing Vectors in Sequences","date":"2022-11-20","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/towards-all-in-one-pre-training-via","title":"Towards All-in-one Pre-training via Maximizing Multi-modal Mutual Information","date":"2022-11-17","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/masked-reconstruction-contrastive-learning","title":"Masked Reconstruction Contrastive Learning with Information Bottleneck Principle","date":"2022-11-15","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/eva-exploring-the-limits-of-masked-visual","title":"EVA: Exploring the Limits of Masked Visual Representation Learning at Scale","date":"2022-11-14","rows_on_this_dataset":1,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":1,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/altclip-altering-the-language-encoder-in-clip","title":"AltCLIP: Altering the Language Encoder in CLIP for Extended Language Capabilities","date":"2022-11-12","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":2,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/internimage-exploring-large-scale-vision","title":"InternImage: Exploring Large-Scale Vision Foundation Models with Deformable Convolutions","date":"2022-11-10","rows_on_this_dataset":6,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":2,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/efficient-multi-order-gated-aggregation","title":"MogaNet: Multi-order Gated Aggregation Network","date":"2022-11-07","rows_on_this_dataset":6,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":12,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/metaformer-baselines-for-vision","title":"MetaFormer Baselines for Vision","date":"2022-10-24","rows_on_this_dataset":32,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":0,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/towards-sustainable-self-supervised-learning","title":"Towards Sustainable Self-supervised Learning","date":"2022-10-20","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":6,"samples_unverified":2,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tokenmixup-efficient-attention-guided-token","title":"TokenMixup: Efficient Attention-guided Token-level Data Augmentation for Transformers","date":"2022-10-14","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/wavemix-lite-a-resource-efficient-neural-1","title":"WaveMix-Lite: A Resource-efficient Neural Network for Image Analysis","date":"2022-10-13","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/online-training-through-time-for-spiking","title":"Online Training Through Time for Spiking Neural Networks","date":"2022-10-09","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tan-without-a-burn-scaling-laws-of-dp-sgd","title":"TAN Without a Burn: Scaling Laws of DP-SGD","date":"2022-10-07","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":0,"samples_unverified":8,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/maple-multi-modal-prompt-learning","title":"MaPLe: Multi-modal Prompt Learning","date":"2022-10-06","rows_on_this_dataset":2,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/moat-alternating-mobile-convolution-and","title":"MOAT: Alternating Mobile Convolution and Attention Brings Strong Vision Models","date":"2022-10-04","rows_on_this_dataset":7,"code_links":2,"syntology":null},{"paper":"/paper/mobilevitv3-mobile-friendly-vision","title":"MobileViTv3: Mobile-Friendly Vision Transformer with Simple and Effective Fusion of Local, Global and Input Features","date":"2022-09-30","rows_on_this_dataset":6,"code_links":2,"syntology":null},{"paper":"/paper/dilated-neighborhood-attention-transformer","title":"Dilated Neighborhood Attention Transformer","date":"2022-09-29","rows_on_this_dataset":8,"code_links":7,"syntology":null},{"paper":"/paper/sapa-similarity-aware-point-affiliation-for","title":"SAPA: Similarity-Aware Point Affiliation for Feature Upsampling","date":"2022-09-26","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/generalized-parametric-contrastive-learning","title":"Generalized Parametric Contrastive Learning","date":"2022-09-26","rows_on_this_dataset":3,"code_links":4,"syntology":null},{"paper":"/paper/mega-moving-average-equipped-gated-attention","title":"Mega: Moving Average Equipped Gated Attention","date":"2022-09-21","rows_on_this_dataset":1,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":13,"samples_ran":12,"samples_unverified":1,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/movq-modulating-quantized-vectors-for-high","title":"MoVQ: Modulating Quantized Vectors for High-Fidelity Image Generation","date":"2022-09-19","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/enhance-the-visual-representation-via","title":"Enhance the Visual Representation via Discrete Adversarial Training","date":"2022-09-16","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":10,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pali-a-jointly-scaled-multilingual-language","title":"PaLI: A Jointly-Scaled Multilingual Language-Image Model","date":"2022-09-14","rows_on_this_dataset":5,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":2,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/exploring-target-representations-for-masked","title":"Exploring Target Representations for Masked Autoencoders","date":"2022-09-08","rows_on_this_dataset":4,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":13,"samples_ran":5,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/gswin-gated-mlp-vision-model-with","title":"gSwin: Gated MLP Vision Model with Hierarchical Structure of Shifted Window","date":"2022-08-24","rows_on_this_dataset":3,"code_links":0,"syntology":null},{"paper":"/paper/semi-supervised-vision-transformers-at-scale","title":"Semi-supervised Vision Transformers at Scale","date":"2022-08-11","rows_on_this_dataset":7,"code_links":1,"syntology":null},{"paper":"/paper/hornet-efficient-high-order-spatial","title":"HorNet: Efficient High-Order Spatial Interactions with Recursive Gated Convolutions","date":"2022-07-28","rows_on_this_dataset":1,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/spatial-channel-token-distillation-for-vision","title":"Spatial-Channel Token Distillation for Vision MLPs","date":"2022-07-23","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/weakly-supervised-object-localization-via","title":"Weakly Supervised Object Localization via Transformer with Implicit Spatial Calibration","date":"2022-07-21","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":7,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tinyvit-fast-pretraining-distillation-for","title":"TinyViT: Fast Pretraining Distillation for Small Vision Transformers","date":"2022-07-21","rows_on_this_dataset":8,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":9,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tailoring-self-supervision-for-supervised","title":"Tailoring Self-Supervision for Supervised Learning","date":"2022-07-20","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/bootstrapped-masked-autoencoders-for-vision","title":"Bootstrapped Masked Autoencoders for Vision BERT Pretraining","date":"2022-07-14","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":6,"samples_unverified":1,"pointer_only_for_licence":7,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unsupervised-visual-representation-learning-4","title":"Unsupervised Visual Representation Learning by Synchronous Momentum Grouping","date":"2022-07-13","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/uninet-unified-architecture-search-with-1","title":"UniNet: Unified Architecture Search with Convolution, Transformer, and MLP","date":"2022-07-12","rows_on_this_dataset":4,"code_links":2,"syntology":null},{"paper":"/paper/next-vit-next-generation-vision-transformer","title":"Next-ViT: Next Generation Vision Transformer for Efficient Deployment in Realistic Industrial Scenarios","date":"2022-07-12","rows_on_this_dataset":3,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":0,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/wave-vit-unifying-wavelet-and-transformers","title":"Wave-ViT: Unifying Wavelet and Transformers for Visual Representation Learning","date":"2022-07-11","rows_on_this_dataset":3,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/np-match-when-neural-processes-meet-semi","title":"NP-Match: When Neural Processes meet Semi-Supervised Learning","date":"2022-07-03","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/revbifpn-the-fully-reversible-bidirectional","title":"RevBiFPN: The Fully Reversible Bidirectional Feature Pyramid Network","date":"2022-06-28","rows_on_this_dataset":7,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vicinity-vision-transformer","title":"Vicinity Vision Transformer","date":"2022-06-21","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":1,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/edgenext-efficiently-amalgamated-cnn","title":"EdgeNeXt: Efficiently Amalgamated CNN-Transformer Architecture for Mobile Vision Applications","date":"2022-06-21","rows_on_this_dataset":2,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":1,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/shapley-nas-discovering-operation-1","title":"Shapley-NAS: Discovering Operation Contribution for Neural Architecture Search","date":"2022-06-20","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/global-context-vision-transformers","title":"Global Context Vision Transformers","date":"2022-06-20","rows_on_this_dataset":5,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":36,"samples_ran":17,"samples_unverified":19,"pointer_only_for_licence":15,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/0-1-deep-neural-networks-via-block-coordinate","title":"0/1 Deep Neural Networks via Block Coordinate Descent","date":"2022-06-19","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/sp-vit-learning-2d-spatial-priors-for-vision","title":"SP-ViT: Learning 2D Spatial Priors for Vision Transformers","date":"2022-06-15","rows_on_this_dataset":6,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":0,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/differentiable-top-k-classification-learning-1","title":"Differentiable Top-k Classification Learning","date":"2022-06-15","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":3,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/an-improved-one-millisecond-mobile-backbone","title":"MobileOne: An Improved One millisecond Mobile Backbone","date":"2022-06-08","rows_on_this_dataset":9,"code_links":10,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":1,"samples_unverified":5,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/separable-self-attention-for-mobile-vision","title":"Separable Self-attention for Mobile Vision Transformers","date":"2022-06-06","rows_on_this_dataset":3,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vision-gnn-an-image-is-worth-graph-of-nodes","title":"Vision GNN: An Image is Worth Graph of Nodes","date":"2022-06-01","rows_on_this_dataset":4,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":34,"samples_ran":19,"samples_unverified":15,"pointer_only_for_licence":22,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/efficientvit-enhanced-linear-attention-for","title":"EfficientViT: Multi-Scale Linear Attention for High-Resolution Dense Prediction","date":"2022-05-29","rows_on_this_dataset":6,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":4,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/wavemix-lite-a-resource-efficient-neural","title":"WaveMix: A Resource-efficient Neural Network for Image Analysis","date":"2022-05-28","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/contrastive-learning-rivals-masked-image","title":"Contrastive Learning Rivals Masked Image Modeling in Fine-tuning via Feature Distillation","date":"2022-05-27","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":2,"samples_unverified":6,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/architecture-agnostic-masked-image-modeling","title":"Architecture-Agnostic Masked Image Modeling -- From ViT back to CNN","date":"2022-05-27","rows_on_this_dataset":8,"code_links":3,"syntology":null},{"paper":"/paper/transboost-improving-the-best-imagenet","title":"TransBoost: Improving the Best ImageNet Performance using Deep Transduction","date":"2022-05-26","rows_on_this_dataset":11,"code_links":1,"syntology":null},{"paper":"/paper/mixmim-mixed-and-masked-image-modeling-for","title":"MixMAE: Mixed and Masked Autoencoder for Efficient Pretraining of Hierarchical Vision Transformers","date":"2022-05-26","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/fast-vision-transformers-with-hilo-attention","title":"Fast Vision Transformers with HiLo Attention","date":"2022-05-26","rows_on_this_dataset":4,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":13,"samples_ran":2,"samples_unverified":11,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/an-evolutionary-approach-to-dynamic","title":"An Evolutionary Approach to Dynamic Introduction of Tasks in Large-scale Multitask Learning Systems","date":"2022-05-25","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/knowledge-distillation-from-a-stronger","title":"Knowledge Distillation from A Stronger Teacher","date":"2022-05-21","rows_on_this_dataset":2,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":10,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deeper-vs-wider-a-revisit-of-transformer","title":"A Study on Transformer Configuration and Training Objective","date":"2022-05-21","rows_on_this_dataset":3,"code_links":0,"syntology":null},{"paper":"/paper/clcnet-rethinking-of-ensemble-modeling-with","title":"CLCNet: Rethinking of Ensemble Modeling with Classification Confidence Network","date":"2022-05-19","rows_on_this_dataset":5,"code_links":1,"syntology":null},{"paper":"/paper/student-collaboration-improves-self","title":"Multiplexed Immunofluorescence Brain Image Analysis Using Self-Supervised Dual-Loss Adaptive Masked Autoencoder","date":"2022-05-10","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/sequencer-deep-lstm-for-image-classification","title":"Sequencer: Deep LSTM for Image Classification","date":"2022-05-04","rows_on_this_dataset":5,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":4,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/coca-contrastive-captioners-are-image-text","title":"CoCa: Contrastive Captioners are Image-Text Foundation Models","date":"2022-05-04","rows_on_this_dataset":3,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":17,"samples_ran":9,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unlocking-high-accuracy-differentially","title":"Unlocking High-Accuracy Differentially Private Image Classification through Scale","date":"2022-04-28","rows_on_this_dataset":1,"code_links":3,"syntology":null},{"paper":"/paper/understanding-the-robustness-in-vision","title":"Understanding The Robustness in Vision Transformers","date":"2022-04-26","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/polyloss-a-polynomial-expansion-perspective-1","title":"PolyLoss: A Polynomial Expansion Perspective of Classification Loss Functions","date":"2022-04-26","rows_on_this_dataset":1,"code_links":19,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":1,"samples_unverified":6,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/gpunet-searching-the-deployable-convolution","title":"GPUNet: Searching the Deployable Convolution Neural Networks for GPUs","date":"2022-04-26","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/adaptive-split-fusion-transformer","title":"Adaptive Split-Fusion Transformer","date":"2022-04-26","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/neighborhood-attention-transformer","title":"Neighborhood Attention Transformer","date":"2022-04-14","rows_on_this_dataset":4,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/minivit-compressing-vision-transformers-with","title":"MiniViT: Compressing Vision Transformers with Weight Multiplexing","date":"2022-04-14","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":0,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/masked-siamese-networks-for-label-efficient","title":"Masked Siamese Networks for Label-Efficient Learning","date":"2022-04-14","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":4,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deit-iii-revenge-of-the-vit","title":"DeiT III: Revenge of the ViT","date":"2022-04-14","rows_on_this_dataset":10,"code_links":12,"syntology":null},{"paper":"/paper/davit-dual-attention-vision-transformers","title":"DaViT: Dual Attention Vision Transformers","date":"2022-04-07","rows_on_this_dataset":8,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":8,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/maxvit-multi-axis-vision-transformer","title":"MaxViT: Multi-Axis Vision Transformer","date":"2022-04-04","rows_on_this_dataset":20,"code_links":15,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":53,"samples_ran":33,"samples_unverified":20,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/revisiting-a-knn-based-image-classification","title":"Revisiting a kNN-based Image Classification System with High-capacity Storage","date":"2022-04-03","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/improving-vision-transformers-by-revisiting","title":"Improving Vision Transformers by Revisiting High-frequency Components","date":"2022-04-03","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":10,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mc-beit-multi-choice-discretization-for-image","title":"mc-BEiT: Multi-choice Discretization for Image BERT Pre-training","date":"2022-03-29","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":3,"samples_unverified":2,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mugs-a-multi-granular-self-supervised","title":"Mugs: A Multi-Granular Self-Supervised Learning Framework","date":"2022-03-27","rows_on_this_dataset":4,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":21,"samples_ran":4,"samples_unverified":17,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deepdpm-deep-clustering-with-an-unknown","title":"DeepDPM: Deep Clustering With an Unknown Number of Clusters","date":"2022-03-27","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/caco-both-positive-and-negative-samples-are","title":"CaCo: Both Positive and Negative Samples are Directly Learnable via Cooperative-adversarial Contrastive Learning","date":"2022-03-27","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":2,"samples_unverified":2,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/three-things-everyone-should-know-about","title":"Three things everyone should know about Vision Transformers","date":"2022-03-18","rows_on_this_dataset":8,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/relational-self-supervised-learning","title":"Weak Augmentation Guided Relational Self-Supervised Learning","date":"2022-03-16","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":0,"samples_unverified":2,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/simmatch-semi-supervised-learning-with","title":"SimMatch: Semi-supervised Learning with Similarity Matching","date":"2022-03-14","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/scaling-up-your-kernels-to-31x31-revisiting","title":"Scaling Up Your Kernels to 31x31: Revisiting Large Kernel Design in CNNs","date":"2022-03-13","rows_on_this_dataset":1,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":2,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deep-autoaugment-1","title":"Deep AutoAugment","date":"2022-03-11","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/activemlp-an-mlp-like-architecture-with","title":"Active Token Mixer","date":"2022-03-11","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/model-soups-averaging-weights-of-multiple","title":"Model soups: averaging weights of multiple fine-tuned models improves accuracy without increasing inference time","date":"2022-03-10","rows_on_this_dataset":4,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":17,"samples_ran":5,"samples_unverified":12,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/conditional-prompt-learning-for-vision","title":"Conditional Prompt Learning for Vision-Language Models","date":"2022-03-10","rows_on_this_dataset":2,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/edgeformer-improving-light-weight-convnets-by","title":"ParC-Net: Position Aware Circular Convolution with Merits from ConvNets and Transformer","date":"2022-03-08","rows_on_this_dataset":1,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":5,"samples_unverified":0,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/b-darts-beta-decay-regularization-for","title":"$β$-DARTS: Beta-Decay Regularization for Differentiable Architecture Search","date":"2022-03-03","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/autoregressive-image-generation-using","title":"Autoregressive Image Generation using Residual Quantization","date":"2022-03-03","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-transformers-for-unsupervised","title":"Self-Supervised Transformers for Unsupervised Object Discovery using Normalized Cut","date":"2022-02-23","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vitaev2-vision-transformer-advanced-by","title":"ViTAEv2: Vision Transformer Advanced by Exploring Inductive Bias for Image Recognition and Beyond","date":"2022-02-21","rows_on_this_dataset":1,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":25,"samples_ran":15,"samples_unverified":10,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/visual-attention-network","title":"Visual Attention Network","date":"2022-02-20","rows_on_this_dataset":9,"code_links":21,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":0,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vision-models-are-more-robust-and-fair-when","title":"Vision Models Are More Robust And Fair When Pretrained On Uncurated Images Without Supervision","date":"2022-02-16","rows_on_this_dataset":6,"code_links":1,"syntology":null},{"paper":"/paper/meta-knowledge-distillation","title":"Meta Knowledge Distillation","date":"2022-02-16","rows_on_this_dataset":4,"code_links":0,"syntology":null},{"paper":"/paper/maskgit-masked-generative-image-transformer","title":"MaskGIT: Masked Generative Image Transformer","date":"2022-02-08","rows_on_this_dataset":1,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":21,"samples_ran":14,"samples_unverified":7,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unifying-architectures-tasks-and-modalities","title":"OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Learning Framework","date":"2022-02-07","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/data2vec-a-general-framework-for-self-1","title":"data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language","date":"2022-02-07","rows_on_this_dataset":1,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":0,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/context-autoencoder-for-self-supervised","title":"Context Autoencoder for Self-Supervised Representation Learning","date":"2022-02-07","rows_on_this_dataset":1,"code_links":6,"syntology":null},{"paper":"/paper/toward-training-at-imagenet-scale-with","title":"Toward Training at ImageNet Scale with Differential Privacy","date":"2022-01-28","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":0,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/when-shift-operation-meets-vision-transformer","title":"When Shift Operation Meets Vision Transformer: An Extremely Simple Alternative to Attention Mechanism","date":"2022-01-26","rows_on_this_dataset":3,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":6,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/uniformer-unifying-convolution-and-self","title":"UniFormer: Unifying Convolution and Self-attention for Visual Recognition","date":"2022-01-24","rows_on_this_dataset":3,"code_links":8,"syntology":null},{"paper":"/paper/patches-are-all-you-need-1","title":"Patches Are All You Need?","date":"2022-01-24","rows_on_this_dataset":1,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":8,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/revisiting-weakly-supervised-pre-training-of","title":"Revisiting Weakly Supervised Pre-Training of Visual Perception Models","date":"2022-01-20","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/omnivore-a-single-model-for-many-visual","title":"Omnivore: A Single Model for Many Visual Modalities","date":"2022-01-20","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pushing-the-limits-of-self-supervised-resnets","title":"Pushing the limits of self-supervised ResNets: Can we outperform supervised learning without labels on ImageNet?","date":"2022-01-13","rows_on_this_dataset":9,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":0,"samples_unverified":14,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-convnet-for-the-2020s","title":"A ConvNet for the 2020s","date":"2022-01-10","rows_on_this_dataset":4,"code_links":54,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":80,"samples_ran":54,"samples_unverified":26,"pointer_only_for_licence":11,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/debiased-learning-from-naturally-imbalanced","title":"Debiased Learning from Naturally Imbalanced Pseudo-Labels","date":"2022-01-05","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vision-transformer-with-deformable-attention","title":"Vision Transformer with Deformable Attention","date":"2022-01-03","rows_on_this_dataset":3,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":8,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-simple-episodic-linear-probe-improves","title":"A Simple Episodic Linear Probe Improves Visual Recognition in the Wild","date":"2022-01-01","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/augmenting-convolutional-networks-with","title":"Augmenting Convolutional networks with attention-based aggregation","date":"2021-12-27","rows_on_this_dataset":7,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/elsa-enhanced-local-self-attention-for-vision","title":"ELSA: Enhanced Local Self-Attention for Vision Transformer","date":"2021-12-23","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/repmlpnet-hierarchical-vision-mlp-with-re","title":"RepMLPNet: Hierarchical Vision MLP with Re-parameterized Locality","date":"2021-12-21","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/max-margin-contrastive-learning","title":"Max-Margin Contrastive Learning","date":"2021-12-21","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":2,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learned-queries-for-efficient-local-attention","title":"Learned Queries for Efficient Local Attention","date":"2021-12-21","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":0,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/masked-feature-prediction-for-self-supervised","title":"Masked Feature Prediction for Self-Supervised Visual Pre-Training","date":"2021-12-16","rows_on_this_dataset":1,"code_links":6,"syntology":null},{"paper":"/paper/nomaro-defending-against-adversarial-attacks","title":"NOMARO: Defending against Adversarial Attacks by NOMA-Inspired Reconstruction Operation","date":"2021-12-14","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/improved-multiscale-vision-transformers-for","title":"MViTv2: Improved Multiscale Vision Transformers for Classification and Detection","date":"2021-12-02","rows_on_this_dataset":5,"code_links":9,"syntology":null},{"paper":"/paper/a-fast-knowledge-distillation-framework-for","title":"A Fast Knowledge Distillation Framework for Visual Recognition","date":"2021-12-02","rows_on_this_dataset":3,"code_links":2,"syntology":null},{"paper":"/paper/information-theoretic-representation","title":"Information Theoretic Representation Distillation","date":"2021-12-01","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/differentiable-spike-rethinking-gradient","title":"Differentiable Spike: Rethinking Gradient-Descent for Training Spiking Neural Networks","date":"2021-12-01","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/boosting-discriminative-visual-representation","title":"Boosting Discriminative Visual Representation Learning with Scenario-Agnostic Mixup","date":"2021-11-30","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/similarity-contrastive-estimation-for-self","title":"Similarity Contrastive Estimation for Self-Supervised Soft Contrastive Learning","date":"2021-11-29","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/fq-vit-fully-quantized-vision-transformer","title":"FQ-ViT: Post-Training Quantization for Fully Quantized Vision Transformer","date":"2021-11-27","rows_on_this_dataset":8,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/peco-perceptual-codebook-for-bert-pre","title":"PeCo: Perceptual Codebook for BERT Pre-training of Vision Transformers","date":"2021-11-24","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/du-darts-decreasing-the-uncertainty-of","title":"DU-DARTS: Decreasing the Uncertainty of Differentiable Architecture Search","date":"2021-11-23","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/semi-supervised-vision-transformers","title":"Semi-Supervised Vision Transformers","date":"2021-11-22","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/metaformer-is-actually-what-you-need-for","title":"MetaFormer Is Actually What You Need for Vision","date":"2021-11-22","rows_on_this_dataset":1,"code_links":18,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/florence-a-new-foundation-model-for-computer","title":"Florence: A New Foundation Model for Computer Vision","date":"2021-11-22","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/discrete-representations-strengthen-vision-1","title":"Discrete Representations Strengthen Vision Transformer Robustness","date":"2021-11-20","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/fbnetv5-neural-architecture-search-for","title":"FBNetV5: Neural Architecture Search for Multiple Tasks in One Run","date":"2021-11-19","rows_on_this_dataset":9,"code_links":0,"syntology":null},{"paper":"/paper/combined-scaling-for-zero-shot-transfer","title":"Combined Scaling for Zero-shot Transfer Learning","date":"2021-11-19","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/swin-transformer-v2-scaling-up-capacity-and","title":"Swin Transformer V2: Scaling Up Capacity and Resolution","date":"2021-11-18","rows_on_this_dataset":4,"code_links":23,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":30,"samples_ran":3,"samples_unverified":27,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/simmim-a-simple-framework-for-masked-image","title":"SimMIM: A Simple Framework for Masked Image Modeling","date":"2021-11-18","rows_on_this_dataset":4,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":9,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/lit-zero-shot-transfer-with-locked-image-text","title":"LiT: Zero-Shot Transfer with Locked-image text Tuning","date":"2021-11-15","rows_on_this_dataset":2,"code_links":5,"syntology":null},{"paper":"/paper/ibot-image-bert-pre-training-with-online","title":"iBOT: Image BERT Pre-Training with Online Tokenizer","date":"2021-11-15","rows_on_this_dataset":9,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/masked-autoencoders-are-scalable-vision","title":"Masked Autoencoders Are Scalable Vision Learners","date":"2021-11-11","rows_on_this_dataset":9,"code_links":58,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":137,"samples_ran":71,"samples_unverified":66,"pointer_only_for_licence":73,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/palette-image-to-image-diffusion-models-1","title":"Palette: Image-to-Image Diffusion Models","date":"2021-11-10","rows_on_this_dataset":6,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/are-transformers-more-robust-than-cnns","title":"Are Transformers More Robust Than CNNs?","date":"2021-11-10","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/sliced-recursive-transformer-1","title":"Sliced Recursive Transformer","date":"2021-11-09","rows_on_this_dataset":5,"code_links":1,"syntology":null},{"paper":"/paper/adaptive-distillation-aggregating-knowledge","title":"Adaptive Distillation: Aggregating Knowledge from Multiple Paths for Efficient Distillation","date":"2021-10-19","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/hrformer-high-resolution-transformer-for","title":"HRFormer: High-Resolution Transformer for Dense Prediction","date":"2021-10-18","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":7,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/flexmatch-boosting-semi-supervised-learning","title":"FlexMatch: Boosting Semi-Supervised Learning with Curriculum Pseudo Labeling","date":"2021-10-15","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-learning-by-estimating-twin-1","title":"Self-Supervised Learning by Estimating Twin Class Distributions","date":"2021-10-14","rows_on_this_dataset":3,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":5,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/weakly-supervised-contrastive-learning-1","title":"Weakly Supervised Contrastive Learning","date":"2021-10-10","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/vector-quantized-image-modeling-with-improved-1","title":"Vector-quantized Image Modeling with Improved VQGAN","date":"2021-10-09","rows_on_this_dataset":1,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":3,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/uninet-unified-architecture-search-with","title":"UniNet: Unified Architecture Search with Convolution, Transformer, and MLP","date":"2021-10-08","rows_on_this_dataset":5,"code_links":0,"syntology":null},{"paper":"/paper/mobilevit-light-weight-general-purpose-and","title":"MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer","date":"2021-10-05","rows_on_this_dataset":2,"code_links":31,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":68,"samples_ran":53,"samples_unverified":15,"pointer_only_for_licence":18,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/resnet-strikes-back-an-improved-training","title":"ResNet strikes back: An improved training procedure in timm","date":"2021-10-01","rows_on_this_dataset":6,"code_links":14,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/nasvit-neural-architecture-search-for","title":"NASViT: Neural Architecture Search for Efficient Vision Transformers with Gradient Conflict aware Supernet Training","date":"2021-09-29","rows_on_this_dataset":13,"code_links":1,"syntology":null},{"paper":"/paper/a-dot-product-attention-free-transformer","title":"A Dot Product Attention Free Transformer","date":"2021-09-29","rows_on_this_dataset":4,"code_links":0,"syntology":null},{"paper":"/paper/compressive-visual-representations","title":"Compressive Visual Representations","date":"2021-09-27","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":0,"samples_unverified":12,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/hptq-hardware-friendly-post-training","title":"HPTQ: Hardware-Friendly Post Training Quantization","date":"2021-09-19","rows_on_this_dataset":5,"code_links":1,"syntology":null},{"paper":"/paper/causal-explanation-of-convolutional-neural","title":"Causal Explanation of Convolutional Neural Networks","date":"2021-09-13","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/sparse-mlp-for-image-recognition-is-self","title":"Sparse MLP for Image Recognition: Is Self-Attention Really Necessary?","date":"2021-09-12","rows_on_this_dataset":3,"code_links":2,"syntology":null},{"paper":"/paper/convmlp-hierarchical-convolutional-mlps-for","title":"ConvMLP: Hierarchical Convolutional MLPs for Vision","date":"2021-09-09","rows_on_this_dataset":3,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/isynet-convolutional-neural-networks-design","title":"ISyNet: Convolutional Neural Networks design for AI accelerator","date":"2021-09-04","rows_on_this_dataset":7,"code_links":9,"syntology":null},{"paper":"/paper/idarts-improving-darts-by-node-normalization","title":"iDARTS: Improving DARTS by Node Normalization and Decorrelation Discretization","date":"2021-08-25","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/contextual-convolutional-neural-networks","title":"Contextual Convolutional Neural Networks","date":"2021-08-17","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/eeea-net-an-early-exit-evolutionary-neural","title":"EEEA-Net: An Early Exit Evolutionary Neural Architecture Search","date":"2021-08-13","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/evo-vit-slow-fast-token-evolution-for-dynamic","title":"Evo-ViT: Slow-Fast Token Evolution for Dynamic Vision Transformer","date":"2021-08-03","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":5,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/group-fisher-pruning-for-practical-network","title":"Group Fisher Pruning for Practical Network Compression","date":"2021-08-02","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/rethinking-and-improving-relative-position","title":"Rethinking and Improving Relative Position Encoding for Vision Transformer","date":"2021-07-29","rows_on_this_dataset":5,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":7,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/hierarchical-self-supervised-augmented","title":"Hierarchical Self-supervised Augmented Knowledge Distillation","date":"2021-07-29","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/parametric-contrastive-learning","title":"Parametric Contrastive Learning","date":"2021-07-26","rows_on_this_dataset":3,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":1,"samples_unverified":2,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/contextual-transformer-networks-for-visual","title":"Contextual Transformer Networks for Visual Recognition","date":"2021-07-26","rows_on_this_dataset":3,"code_links":7,"syntology":null},{"paper":"/paper/go-wider-instead-of-deeper","title":"Go Wider Instead of Deeper","date":"2021-07-25","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/m-darts-model-uncertainty-aware","title":"$μ$DARTS: Model Uncertainty-Aware Differentiable Architecture Search","date":"2021-07-24","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/bias-loss-for-mobile-neural-networks","title":"Bias Loss for Mobile Neural Networks","date":"2021-07-23","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/cyclemlp-a-mlp-like-architecture-for-dense","title":"CycleMLP: A MLP-like Architecture for Dense Prediction","date":"2021-07-21","rows_on_this_dataset":1,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":8,"samples_unverified":7,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/ressl-relational-self-supervised-learning","title":"ReSSL: Relational Self-Supervised Learning with Weak Augmentation","date":"2021-07-20","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/visual-parser-representing-part-whole","title":"Visual Parser: Representing Part-whole Hierarchies with Transformers","date":"2021-07-13","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/collaboration-of-experts-achieving-80-top-1","title":"Collaboration of Experts: Achieving 80% Top-1 Accuracy on ImageNet with 100M FLOPs","date":"2021-07-08","rows_on_this_dataset":3,"code_links":0,"syntology":null},{"paper":"/paper/glit-neural-architecture-search-for-global","title":"GLiT: Neural Architecture Search for Global and Local Image Transformer","date":"2021-07-07","rows_on_this_dataset":3,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":2,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/global-filter-networks-for-image","title":"Global Filter Networks for Image Classification","date":"2021-07-01","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":6,"samples_unverified":4,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/cswin-transformer-a-general-vision","title":"CSWin Transformer: A General Vision Transformer Backbone with Cross-Shaped Windows","date":"2021-07-01","rows_on_this_dataset":1,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":2,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/autoformer-searching-transformers-for-visual","title":"AutoFormer: Searching Transformers for Visual Recognition","date":"2021-07-01","rows_on_this_dataset":3,"code_links":2,"syntology":null},{"paper":"/paper/pvtv2-improved-baselines-with-pyramid-vision","title":"PVT v2: Improved Baselines with Pyramid Vision Transformer","date":"2021-06-25","rows_on_this_dataset":5,"code_links":18,"syntology":null},{"paper":"/paper/volo-vision-outlooker-for-visual-recognition","title":"VOLO: Vision Outlooker for Visual Recognition","date":"2021-06-24","rows_on_this_dataset":7,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":1,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/ac-dc-alternating-compressed-decompressed","title":"AC/DC: Alternating Compressed/DeCompressed Training of Deep Neural Networks","date":"2021-06-23","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":2,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tokenlearner-what-can-8-learned-tokens-do-for","title":"TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?","date":"2021-06-21","rows_on_this_dataset":2,"code_links":11,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/sparse-training-via-boosting-pruning","title":"Sparse Training via Boosting Pruning Plasticity with Neuroregeneration","date":"2021-06-19","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/xcit-cross-covariance-image-transformers","title":"XCiT: Cross-Covariance Image Transformers","date":"2021-06-17","rows_on_this_dataset":4,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":3,"samples_unverified":11,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/efficient-self-supervised-vision-transformers","title":"Efficient Self-supervised Vision Transformers for Representation Learning","date":"2021-06-17","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":2,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/how-does-topology-of-neural-architectures","title":"How does topology of neural architectures impact gradient propagation and model performance?","date":"2021-06-16","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/beit-bert-pre-training-of-image-transformers","title":"BEiT: BERT Pre-Training of Image Transformers","date":"2021-06-15","rows_on_this_dataset":4,"code_links":14,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":6,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/scaling-vision-with-sparse-mixture-of-experts","title":"Scaling Vision with Sparse Mixture of Experts","date":"2021-06-10","rows_on_this_dataset":17,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/knowledge-distillation-a-good-teacher-is","title":"Knowledge distillation: A good teacher is patient and consistent","date":"2021-06-09","rows_on_this_dataset":1,"code_links":10,"syntology":null},{"paper":"/paper/coatnet-marrying-convolution-and-attention","title":"CoAtNet: Marrying Convolution and Attention for All Data Sizes","date":"2021-06-09","rows_on_this_dataset":7,"code_links":14,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":2,"samples_unverified":3,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/scaling-vision-transformers","title":"Scaling Vision Transformers","date":"2021-06-08","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/vitae-vision-transformer-advanced-by","title":"ViTAE: Vision Transformer Advanced by Exploring Intrinsic Inductive Bias","date":"2021-06-07","rows_on_this_dataset":6,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/refiner-refining-self-attention-for-vision","title":"Refiner: Refining Self-attention for Vision Transformers","date":"2021-06-07","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/x-volution-on-the-unification-of-convolution","title":"X-volution: On the unification of convolution and self-attention","date":"2021-06-04","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/when-vision-transformers-outperform-resnets","title":"When Vision Transformers Outperform ResNets without Pre-training or Strong Data Augmentations","date":"2021-06-03","rows_on_this_dataset":6,"code_links":2,"syntology":null},{"paper":"/paper/dynamicvit-efficient-vision-transformers-with","title":"DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification","date":"2021-06-03","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":6,"samples_unverified":3,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/container-context-aggregation-network","title":"Container: Context Aggregation Network","date":"2021-06-02","rows_on_this_dataset":2,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":0,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/not-all-images-are-worth-16x16-words-dynamic","title":"Not All Images are Worth 16x16 Words: Dynamic Transformers for Efficient Image Recognition","date":"2021-05-31","rows_on_this_dataset":3,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/rest-an-efficient-transformer-for-visual","title":"ResT: An Efficient Transformer for Visual Recognition","date":"2021-05-28","rows_on_this_dataset":2,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":1,"samples_unverified":8,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/drawing-multiple-augmentation-samples-per","title":"Drawing Multiple Augmentation Samples Per Image During Training Efficiently Decreases Test Error","date":"2021-05-27","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/aggregating-nested-transformers","title":"Nested Hierarchical Transformer: Towards Accurate, Data-Efficient and Interpretable Visual Understanding","date":"2021-05-26","rows_on_this_dataset":3,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":26,"samples_ran":7,"samples_unverified":19,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unsupervised-visual-representation-learning-3","title":"Unsupervised Visual Representation Learning by Online Constrained K-Means","date":"2021-05-24","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":3,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/correlated-input-dependent-label-noise-in","title":"Correlated Input-Dependent Label Noise in Large-Scale Image Classification","date":"2021-05-19","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/rethinking-the-design-principles-of-robust","title":"Towards Robust Vision Transformer","date":"2021-05-17","rows_on_this_dataset":3,"code_links":2,"syntology":null},{"paper":"/paper/pay-attention-to-mlps","title":"Pay Attention to MLPs","date":"2021-05-17","rows_on_this_dataset":1,"code_links":20,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":44,"samples_ran":34,"samples_unverified":10,"pointer_only_for_licence":11,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/divide-and-contrast-self-supervised-learning","title":"Divide and Contrast: Self-supervised Learning from Uncurated Data","date":"2021-05-17","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/vicreg-variance-invariance-covariance","title":"VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learning","date":"2021-05-11","rows_on_this_dataset":2,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":17,"samples_ran":11,"samples_unverified":6,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-learning-with-swin","title":"Self-Supervised Learning with Swin Transformers","date":"2021-05-10","rows_on_this_dataset":2,"code_links":6,"syntology":null},{"paper":"/paper/conformer-local-features-coupling-global","title":"Conformer: Local Features Coupling Global Representations for Visual Recognition","date":"2021-05-09","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":7,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/resmlp-feedforward-networks-for-image","title":"ResMLP: Feedforward networks for image classification with data-efficient training","date":"2021-05-07","rows_on_this_dataset":12,"code_links":19,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":2,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/network-pruning-that-matters-a-case-study-on-1","title":"Network Pruning That Matters: A Case Study on Retraining Variants","date":"2021-05-07","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":2,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/basisnet-two-stage-model-synthesis-for-1","title":"BasisNet: Two-stage Model Synthesis for Efficient Inference","date":"2021-05-07","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/do-you-even-need-attention-a-stack-of-feed","title":"Do You Even Need Attention? A Stack of Feed-Forward Layers Does Surprisingly Well on ImageNet","date":"2021-05-06","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/repmlp-re-parameterizing-convolutions-into","title":"RepMLP: Re-parameterizing Convolutions into Fully-connected Layers for Image Recognition","date":"2021-05-05","rows_on_this_dataset":1,"code_links":10,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":2,"samples_unverified":12,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/beyond-self-attention-external-attention","title":"Beyond Self-attention: External Attention using Two Linear Layers for Visual Tasks","date":"2021-05-05","rows_on_this_dataset":1,"code_links":7,"syntology":null},{"paper":"/paper/mlp-mixer-an-all-mlp-architecture-for-vision","title":"MLP-Mixer: An all-MLP Architecture for Vision","date":"2021-05-04","rows_on_this_dataset":3,"code_links":49,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":134,"samples_ran":106,"samples_unverified":28,"pointer_only_for_licence":36,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/with-a-little-help-from-my-friends-nearest","title":"With a Little Help from My Friends: Nearest-Neighbor Contrastive Learning of Visual Representations","date":"2021-04-29","rows_on_this_dataset":3,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/emerging-properties-in-self-supervised-vision","title":"Emerging Properties in Self-Supervised Vision Transformers","date":"2021-04-29","rows_on_this_dataset":7,"code_links":32,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":20,"samples_ran":5,"samples_unverified":15,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/twins-revisiting-spatial-attention-design-in","title":"Twins: Revisiting the Design of Spatial Attention in Vision Transformers","date":"2021-04-28","rows_on_this_dataset":1,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":0,"samples_unverified":2,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/semi-supervised-learning-of-visual-features","title":"Semi-Supervised Learning of Visual Features by Non-Parametrically Predicting View Assignments with Support Samples","date":"2021-04-28","rows_on_this_dataset":8,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":21,"samples_ran":7,"samples_unverified":14,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/visformer-the-vision-friendly-transformer","title":"Visformer: The Vision-friendly Transformer","date":"2021-04-26","rows_on_this_dataset":2,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/token-labeling-training-a-85-5-top-1-accuracy","title":"All Tokens Matter: Token Labeling for Training Better Vision Transformers","date":"2021-04-22","rows_on_this_dataset":3,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":3,"samples_unverified":2,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/multiscale-vision-transformers","title":"Multiscale Vision Transformers","date":"2021-04-22","rows_on_this_dataset":2,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":26,"samples_ran":13,"samples_unverified":13,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/differentiable-model-compression-via-pseudo","title":"Differentiable Model Compression via Pseudo Quantization Noise","date":"2021-04-20","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/distilling-knowledge-via-knowledge-review","title":"Distilling Knowledge via Knowledge Review","date":"2021-04-19","rows_on_this_dataset":1,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":5,"samples_unverified":3,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/solving-inefficiency-of-self-supervised","title":"Solving Inefficiency of Self-supervised Representation Learning","date":"2021-04-18","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":0,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/polynomial-networks-in-deep-classifiers","title":"Augmenting Deep Classifiers with Polynomial Neural Networks","date":"2021-04-16","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/bnn-bn-training-binary-neural-networks","title":"\"BNN - BN = ?\": Training Binary Neural Networks without Batch Normalization","date":"2021-04-16","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":4,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/asymmnet-towards-ultralight-convolution","title":"AsymmNet: Towards ultralight convolution neural networks using asymmetrical bottlenecks","date":"2021-04-15","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/localvit-bringing-locality-to-vision","title":"LocalViT: Bringing Locality to Vision Transformers","date":"2021-04-12","rows_on_this_dataset":5,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":3,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/escaping-the-big-data-paradigm-with-compact","title":"Escaping the Big Data Paradigm with Compact Transformers","date":"2021-04-12","rows_on_this_dataset":1,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":3,"samples_unverified":3,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/an-empirical-study-of-training-self","title":"An Empirical Study of Training Self-Supervised Vision Transformers","date":"2021-04-05","rows_on_this_dataset":7,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/levit-a-vision-transformer-in-convnet-s","title":"LeViT: a Vision Transformer in ConvNet's Clothing for Faster Inference","date":"2021-04-02","rows_on_this_dataset":10,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":30,"samples_ran":23,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/training-multi-bit-quantized-and-binarized","title":"Training Multi-bit Quantized and Binarized Networks with A Learnable Symmetric Quantizer","date":"2021-04-01","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/efficientnetv2-smaller-models-and-faster","title":"EfficientNetV2: Smaller Models and Faster Training","date":"2021-04-01","rows_on_this_dataset":7,"code_links":26,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":79,"samples_ran":41,"samples_unverified":38,"pointer_only_for_licence":10,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/going-deeper-with-image-transformers","title":"Going deeper with Image Transformers","date":"2021-03-31","rows_on_this_dataset":12,"code_links":21,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":5,"samples_unverified":6,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/rethinking-spatial-dimensions-of-vision","title":"Rethinking Spatial Dimensions of Vision Transformers","date":"2021-03-30","rows_on_this_dataset":4,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":20,"samples_ran":10,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/cvt-introducing-convolutions-to-vision","title":"CvT: Introducing Convolutions to Vision Transformers","date":"2021-03-29","rows_on_this_dataset":7,"code_links":16,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":47,"samples_ran":29,"samples_unverified":18,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/complementary-relation-contrastive","title":"Complementary Relation Contrastive Distillation","date":"2021-03-29","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/2103-15358","title":"Multi-Scale Vision Longformer: A New Vision Transformer for High-Resolution Image Encoding","date":"2021-03-29","rows_on_this_dataset":6,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":6,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/2103-14899","title":"CrossViT: Cross-Attention Multi-Scale Vision Transformer for Image Classification","date":"2021-03-27","rows_on_this_dataset":4,"code_links":15,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":26,"samples_ran":17,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/swin-transformer-hierarchical-vision","title":"Swin Transformer: Hierarchical Vision Transformer using Shifted Windows","date":"2021-03-25","rows_on_this_dataset":3,"code_links":80,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":207,"samples_ran":108,"samples_unverified":99,"pointer_only_for_licence":43,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/automix-unveiling-the-power-of-mixup","title":"AutoMix: Unveiling the Power of Mixup for Stronger Classifiers","date":"2021-03-24","rows_on_this_dataset":4,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/scaling-local-self-attention-for-parameter","title":"Scaling Local Self-Attention for Parameter Efficient Visual Backbones","date":"2021-03-23","rows_on_this_dataset":1,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":20,"samples_ran":12,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/bossnas-exploring-hybrid-cnn-transformers","title":"BossNAS: Exploring Hybrid CNN-transformers with Block-wisely Self-supervised Neural Architecture Search","date":"2021-03-23","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/incorporating-convolution-designs-into-visual","title":"Incorporating Convolution Designs into Visual Transformers","date":"2021-03-22","rows_on_this_dataset":4,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":9,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deepvit-towards-deeper-vision-transformer","title":"DeepViT: Towards Deeper Vision Transformer","date":"2021-03-22","rows_on_this_dataset":2,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-classification-network","title":"Self-Supervised Classification Network","date":"2021-03-19","rows_on_this_dataset":2,"code_links":2,"syntology":null},{"paper":"/paper/scalable-visual-transformers-with","title":"Scalable Vision Transformers with Hierarchical Pooling","date":"2021-03-19","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/convit-improving-vision-transformers-with","title":"ConViT: Improving Vision Transformers with Soft Convolutional Inductive Biases","date":"2021-03-19","rows_on_this_dataset":6,"code_links":9,"syntology":null},{"paper":"/paper/trivialaugment-tuning-free-yet-state-of-the","title":"TrivialAugment: Tuning-free Yet State-of-the-Art Data Augmentation","date":"2021-03-18","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/multi-prize-lottery-ticket-hypothesis-finding-1","title":"Multi-Prize Lottery Ticket Hypothesis: Finding Accurate Binary Neural Networks by Pruning A Randomly Weighted Network","date":"2021-03-17","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/revisiting-resnets-improved-training-and","title":"Revisiting ResNets: Improved Training and Scaling Strategies","date":"2021-03-13","rows_on_this_dataset":2,"code_links":3,"syntology":null},{"paper":"/paper/involution-inverting-the-inherence-of","title":"Involution: Inverting the Inherence of Convolution for Visual Recognition","date":"2021-03-10","rows_on_this_dataset":5,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":7,"samples_unverified":2,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/perceiver-general-perception-with-iterative","title":"Perceiver: General Perception with Iterative Attention","date":"2021-03-04","rows_on_this_dataset":2,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":55,"samples_ran":41,"samples_unverified":14,"pointer_only_for_licence":11,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/barlow-twins-self-supervised-learning-via","title":"Barlow Twins: Self-Supervised Learning via Redundancy Reduction","date":"2021-03-04","rows_on_this_dataset":3,"code_links":24,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":26,"samples_ran":21,"samples_unverified":5,"pointer_only_for_licence":10,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-pretraining-of-visual","title":"Self-supervised Pretraining of Visual Features in the Wild","date":"2021-03-02","rows_on_this_dataset":7,"code_links":1,"syntology":null},{"paper":"/paper/transformer-in-transformer","title":"Transformer in Transformer","date":"2021-02-27","rows_on_this_dataset":1,"code_links":12,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":24,"samples_ran":16,"samples_unverified":8,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learning-transferable-visual-models-from","title":"Learning Transferable Visual Models From Natural Language Supervision","date":"2021-02-26","rows_on_this_dataset":9,"code_links":82,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":20,"samples_ran":16,"samples_unverified":4,"pointer_only_for_licence":16,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/hardcore-nas-hard-constrained-differentiable","title":"HardCoRe-NAS: Hard Constrained diffeRentiable Neural Architecture Search","date":"2021-02-23","rows_on_this_dataset":11,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":2,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/lambdanetworks-modeling-long-range-1","title":"LambdaNetworks: Modeling Long-Range Interactions Without Attention","date":"2021-02-17","rows_on_this_dataset":2,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/centroid-transformers-learning-to-abstract","title":"Centroid Transformers: Learning to Abstract with Attention","date":"2021-02-17","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/alphanet-improved-training-of-supernet-with","title":"AlphaNet: Improved Training of Supernets with Alpha-Divergence","date":"2021-02-16","rows_on_this_dataset":15,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":31,"samples_ran":16,"samples_unverified":15,"pointer_only_for_licence":31,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-large-batch-optimizer-reality-check","title":"A Large Batch Optimizer Reality Check: Traditional, Generic Optimizers Suffice Across Batch Sizes","date":"2021-02-12","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/scaling-up-visual-and-vision-language","title":"Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision","date":"2021-02-11","rows_on_this_dataset":3,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":8,"samples_unverified":2,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/high-performance-large-scale-image","title":"High-Performance Large-Scale Image Recognition Without Normalization","date":"2021-02-11","rows_on_this_dataset":9,"code_links":20,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/show-attend-and-distill-knowledge","title":"Show, Attend and Distill:Knowledge Distillation via Attention-based Feature Matching","date":"2021-02-05","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/do-we-actually-need-dense-over","title":"Do We Actually Need Dense Over-Parameterization? In-Time Over-Parameterization in Sparse Training","date":"2021-02-04","rows_on_this_dataset":2,"code_links":4,"syntology":null},{"paper":"/paper/zen-nas-a-zero-shot-nas-for-high-performance","title":"Zen-NAS: A Zero-Shot NAS for High-Performance Deep Image Recognition","date":"2021-02-01","rows_on_this_dataset":4,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/rethinking-soft-labels-for-knowledge-1","title":"Rethinking Soft Labels for Knowledge Distillation: A Bias-Variance Tradeoff Perspective","date":"2021-02-01","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tokens-to-token-vit-training-vision","title":"Tokens-to-Token ViT: Training Vision Transformers from Scratch on ImageNet","date":"2021-01-28","rows_on_this_dataset":6,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":26,"samples_ran":21,"samples_unverified":5,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/bottleneck-transformers-for-visual","title":"Bottleneck Transformers for Visual Recognition","date":"2021-01-27","rows_on_this_dataset":12,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":49,"samples_ran":26,"samples_unverified":23,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/exponential-moving-average-normalization-for","title":"Exponential Moving Average Normalization for Self-supervised and Semi-supervised Learning","date":"2021-01-21","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/re-labeling-imagenet-from-single-to-multi","title":"Re-labeling ImageNet: from Single to Multi-Labels, from Global to Localized Labels","date":"2021-01-13","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/repvgg-making-vgg-style-convnets-great-again","title":"RepVGG: Making VGG-style ConvNets Great Again","date":"2021-01-11","rows_on_this_dataset":2,"code_links":25,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":16,"samples_ran":13,"samples_unverified":3,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/contextual-classification-using-self","title":"Contextual Classification Using Self-Supervised Auxiliary Models for Deep Neural Networks","date":"2021-01-07","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/autodropout-learning-dropout-patterns-to","title":"AutoDropout: Learning Dropout Patterns to Regularize Deep Networks","date":"2021-01-05","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/sedona-search-for-decoupled-neural-networks","title":"SEDONA: Search for Decoupled Neural Networks toward Greedy Block-wise Learning","date":"2021-01-01","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/exploring-inter-channel-correlation-for","title":"Exploring Inter-Channel Correlation for Diversity-Preserved Knowledge Distillation","date":"2021-01-01","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/distilling-global-and-local-logits-with","title":"Distilling Global and Local Logits With Densely Connected Relations","date":"2021-01-01","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/self-supervised-pre-training-with-hard","title":"Self-supervised Pre-training with Hard Examples Improves Visual Representations","date":"2020-12-25","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/training-data-efficient-image-transformers","title":"Training data-efficient image transformers & distillation through attention","date":"2020-12-23","rows_on_this_dataset":4,"code_links":40,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":19,"samples_ran":12,"samples_unverified":7,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/evolving-neural-architecture-using-one-shot","title":"Evolving Neural Architecture Using One Shot Model","date":"2020-12-23","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/online-bag-of-visual-words-generation-for","title":"OBoW: Online Bag-of-Visual-Words Generation for Self-Supervised Learning","date":"2020-12-21","rows_on_this_dataset":3,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/taming-transformers-for-high-resolution-image","title":"Taming Transformers for High-Resolution Image Synthesis","date":"2020-12-17","rows_on_this_dataset":1,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":6,"samples_unverified":0,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/splitnet-divide-and-co-training","title":"Towards Better Accuracy-efficiency Trade-offs: Divide and Co-training","date":"2020-11-30","rows_on_this_dataset":3,"code_links":2,"syntology":null},{"paper":"/paper/torchdistill-a-modular-configuration-driven","title":"torchdistill: A Modular, Configuration-Driven Framework for Knowledge Distillation","date":"2020-11-25","rows_on_this_dataset":7,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":26,"samples_ran":10,"samples_unverified":16,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/grafit-learning-fine-grained-image","title":"Grafit: Learning fine-grained image representations with coarse labels","date":"2020-11-25","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/comatch-semi-supervised-learning-with","title":"CoMatch: Semi-supervised Learning with Contrastive Graph Regularization","date":"2020-11-23","rows_on_this_dataset":2,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":2,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/boosting-contrastive-self-supervised-learning","title":"Boosting Contrastive Self-Supervised Learning with False Negative Cancellation","date":"2020-11-23","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/exploring-simple-siamese-representation","title":"Exploring Simple Siamese Representation Learning","date":"2020-11-20","rows_on_this_dataset":1,"code_links":26,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":39,"samples_ran":34,"samples_unverified":5,"pointer_only_for_licence":22,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/attentivenas-improving-neural-architecture","title":"AttentiveNAS: Improving Neural Architecture Search via Attentive Sampling","date":"2020-11-18","rows_on_this_dataset":6,"code_links":2,"syntology":null},{"paper":"/paper/learning-visual-representations-for-transfer-1","title":"Learning Visual Representations for Transfer Learning by Suppressing Texture","date":"2020-11-03","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/in-defense-of-feature-mimicking-for-knowledge","title":"Distilling Knowledge by Mimicking Features","date":"2020-11-03","rows_on_this_dataset":1,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":5,"samples_unverified":1,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/model-rubik-s-cube-twisting-resolution-depth","title":"Model Rubik's Cube: Twisting Resolution, Depth and Width for TinyNets","date":"2020-10-28","rows_on_this_dataset":2,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/an-image-is-worth-16x16-words-transformers-1","title":"An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale","date":"2020-10-22","rows_on_this_dataset":3,"code_links":158,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":419,"samples_ran":281,"samples_unverified":138,"pointer_only_for_licence":154,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/representation-learning-via-invariant-causal-1","title":"Representation Learning via Invariant Causal Mechanisms","date":"2020-10-15","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/shape-texture-debiased-neural-network-1","title":"Shape-Texture Debiased Neural Network Training","date":"2020-10-12","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":1,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/sharpness-aware-minimization-for-efficiently-1","title":"Sharpness-Aware Minimization for Efficiently Improving Generalization","date":"2020-10-03","rows_on_this_dataset":2,"code_links":18,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":20,"samples_ran":8,"samples_unverified":12,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/attentional-feature-fusion","title":"Attentional Feature Fusion","date":"2020-09-29","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/neural-architecture-search-using-stable-rank","title":"MSR-DARTS: Minimum Stable Rank of Differentiable Architecture Search","date":"2020-09-19","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/meal-v2-boosting-vanilla-resnet-50-to-80-top","title":"MEAL V2: Boosting Vanilla ResNet-50 to 80%+ Top-1 Accuracy on ImageNet without Tricks","date":"2020-09-17","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/puzzle-mix-exploiting-saliency-and-local-1","title":"Puzzle Mix: Exploiting Saliency and Local Statistics for Optimal Mixup","date":"2020-09-15","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/quantnet-learning-to-quantize-by-learning","title":"QuantNet: Learning to Quantize by Learning within Fully Differentiable Framework","date":"2020-09-10","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/self-supervised-learning-for-large-scale","title":"Self-Supervised Learning for Large-Scale Unsupervised Image Clustering","date":"2020-08-24","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":2,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/nsganetv2-evolutionary-multi-objective","title":"NSGANetV2: Evolutionary Multi-Objective Surrogate-Assisted Neural Architecture Search","date":"2020-07-20","rows_on_this_dataset":4,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":2,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/hmq-hardware-friendly-mixed-precision","title":"HMQ: Hardware Friendly Mixed Precision Quantization Block for CNNs","date":"2020-07-20","rows_on_this_dataset":4,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":0,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/generative-pretraining-from-pixels","title":"Generative Pretraining from Pixels","date":"2020-07-17","rows_on_this_dataset":4,"code_links":4,"syntology":null},{"paper":"/paper/neural-architecture-search-with-gbdt","title":"Accuracy Prediction with Non-neural Model for Neural Architecture Search","date":"2020-07-09","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":0,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/eagleeye-fast-sub-net-evaluation-for","title":"EagleEye: Fast Sub-net Evaluation for Efficient Neural Network Pruning","date":"2020-07-06","rows_on_this_dataset":5,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/rexnet-diminishing-representational","title":"Rethinking Channel Dimensions for Efficient Model Design","date":"2020-07-02","rows_on_this_dataset":9,"code_links":10,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":1,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-knowledge-distillation-a-simple-way-for","title":"Self-Knowledge Distillation with Progressive Refinement of Targets","date":"2020-06-22","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pyramidal-convolution-rethinking","title":"Pyramidal Convolution: Rethinking Convolutional Neural Networks for Visual Recognition","date":"2020-06-20","rows_on_this_dataset":1,"code_links":3,"syntology":null},{"paper":"/paper/deep-polynomial-neural-networks","title":"Deep Polynomial Neural Networks","date":"2020-06-20","rows_on_this_dataset":1,"code_links":5,"syntology":null},{"paper":"/paper/semi-supervised-recognition-under-a-noisy-and","title":"Semi-Supervised Recognition under a Noisy and Fine-grained Dataset","date":"2020-06-18","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/drnas-dirichlet-neural-architecture-search","title":"DrNAS: Dirichlet Neural Architecture Search","date":"2020-06-18","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/unsupervised-learning-of-visual-features-by","title":"Unsupervised Learning of Visual Features by Contrasting Cluster Assignments","date":"2020-06-17","rows_on_this_dataset":7,"code_links":18,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":17,"samples_ran":13,"samples_unverified":4,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/big-self-supervised-models-are-strong-semi","title":"Big Self-Supervised Models are Strong Semi-Supervised Learners","date":"2020-06-17","rows_on_this_dataset":16,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":0,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/multiscale-deep-equilibrium-models","title":"Multiscale Deep Equilibrium Models","date":"2020-06-15","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":10,"samples_unverified":4,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/bootstrap-your-own-latent-a-new-approach-to","title":"Bootstrap your own latent: A new approach to self-supervised Learning","date":"2020-06-13","rows_on_this_dataset":8,"code_links":31,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":79,"samples_ran":62,"samples_unverified":17,"pointer_only_for_licence":46,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/knowledge-distillation-meets-self-supervision","title":"Knowledge Distillation Meets Self-Supervision","date":"2020-06-12","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fbnetv3-joint-architecture-recipe-search","title":"FBNetV3: Joint Architecture-Recipe Search using Predictor Pretraining","date":"2020-06-03","rows_on_this_dataset":4,"code_links":2,"syntology":null},{"paper":"/paper/memnas-memory-efficient-neural-architecture","title":"MemNAS: Memory-Efficient Neural Architecture Search With Grow-Trim Learning","date":"2020-06-01","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/learning-to-classify-images-without-labels","title":"SCAN: Learning to Classify Images without Labels","date":"2020-05-25","rows_on_this_dataset":3,"code_links":2,"syntology":null},{"paper":"/paper/what-makes-for-good-views-for-contrastive","title":"What Makes for Good Views for Contrastive Learning?","date":"2020-05-20","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/optimizing-neural-architecture-search-using","title":"Optimizing Neural Architecture Search using Limited GPU Time in a Dynamic Search Space: A Gene Expression Programming Approach","date":"2020-05-15","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/neural-architecture-transfer","title":"Neural Architecture Transfer","date":"2020-05-12","rows_on_this_dataset":5,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/prototypical-contrastive-learning-of","title":"Prototypical Contrastive Learning of Unsupervised Representations","date":"2020-05-11","rows_on_this_dataset":4,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":2,"samples_unverified":3,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/noisy-differentiable-architecture-search","title":"Noisy Differentiable Architecture Search","date":"2020-05-07","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/supervised-contrastive-learning","title":"Supervised Contrastive Learning","date":"2020-04-23","rows_on_this_dataset":1,"code_links":26,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":23,"samples_ran":6,"samples_unverified":17,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/lsq-improving-low-bit-quantization-through","title":"LSQ+: Improving low-bit quantization through learnable offsets and better initialization","date":"2020-04-20","rows_on_this_dataset":2,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":3,"samples_unverified":7,"pointer_only_for_licence":10,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/resnest-split-attention-networks","title":"ResNeSt: Split-Attention Networks","date":"2020-04-19","rows_on_this_dataset":5,"code_links":36,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":48,"samples_ran":8,"samples_unverified":40,"pointer_only_for_licence":23,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/geometry-aware-gradient-algorithms-for-neural","title":"Geometry-Aware Gradient Algorithms for Neural Architecture Search","date":"2020-04-16","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":12,"samples_unverified":3,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fbnetv2-differentiable-neural-architecture","title":"FBNetV2: Differentiable Neural Architecture Search for Spatial and Channel Dimensions","date":"2020-04-12","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/neural-architecture-search-for-lightweight","title":"Neural Architecture Search for Lightweight Non-Local Networks","date":"2020-04-04","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-generic-graph-based-neural-architecture","title":"A Generic Graph-based Neural Architecture Encoding Scheme for Predictor-based NAS","date":"2020-04-04","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/uniformaugment-a-search-free-probabilistic","title":"UniformAugment: A Search-free Probabilistic Data Augmentation Approach","date":"2020-03-31","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/muxconv-information-multiplexing-in","title":"MUXConv: Information Multiplexing in Convolutional Neural Networks","date":"2020-03-31","rows_on_this_dataset":8,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":3,"samples_unverified":2,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/tresnet-high-performance-gpu-dedicated","title":"TResNet: High Performance GPU-Dedicated Architecture","date":"2020-03-30","rows_on_this_dataset":1,"code_links":3,"syntology":null},{"paper":"/paper/designing-network-design-spaces","title":"Designing Network Design Spaces","date":"2020-03-30","rows_on_this_dataset":6,"code_links":26,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":53,"samples_ran":11,"samples_unverified":42,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/milking-cowmask-for-semi-supervised-image","title":"Milking CowMask for Semi-Supervised Image Classification","date":"2020-03-26","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":1,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/greedynas-towards-fast-one-shot-nas-with","title":"GreedyNAS: Towards Fast One-Shot NAS with Greedy Supernet","date":"2020-03-25","rows_on_this_dataset":6,"code_links":0,"syntology":null},{"paper":"/paper/circumventing-outliers-of-autoaugment-with","title":"Circumventing Outliers of AutoAugment with Knowledge Distillation","date":"2020-03-25","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/bignas-scaling-up-neural-architecture-search","title":"BigNAS: Scaling Up Neural Architecture Search with Big Single-Stage Models","date":"2020-03-24","rows_on_this_dataset":3,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":0,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/meta-pseudo-labels","title":"Meta Pseudo Labels","date":"2020-03-23","rows_on_this_dataset":4,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":14,"samples_ran":5,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fixing-the-train-test-resolution-discrepancy-2","title":"Fixing the train-test resolution discrepancy: FixEfficientNet","date":"2020-03-18","rows_on_this_dataset":11,"code_links":1,"syntology":null},{"paper":"/paper/improved-baselines-with-momentum-contrastive","title":"Improved Baselines with Momentum Contrastive Learning","date":"2020-03-09","rows_on_this_dataset":2,"code_links":36,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":43,"samples_ran":8,"samples_unverified":35,"pointer_only_for_licence":10,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/dada-differentiable-automatic-data","title":"DADA: Differentiable Automatic Data Augmentation","date":"2020-03-08","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":5,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/semi-supervised-neural-architecture-search","title":"Semi-Supervised Neural Architecture Search","date":"2020-02-24","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":3,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/maxup-a-simple-way-to-improve-generalization","title":"MaxUp: A Simple Way to Improve Generalization of Neural Network Training","date":"2020-02-20","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/knapsack-pruning-with-inner-distillation","title":"Knapsack Pruning with Inner Distillation","date":"2020-02-19","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-simple-framework-for-contrastive-learning","title":"A Simple Framework for Contrastive Learning of Visual Representations","date":"2020-02-13","rows_on_this_dataset":11,"code_links":96,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":137,"samples_ran":79,"samples_unverified":58,"pointer_only_for_licence":52,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fixmatch-simplifying-semi-supervised-learning","title":"FixMatch: Simplifying Semi-Supervised Learning with Consistency and Confidence","date":"2020-01-21","rows_on_this_dataset":1,"code_links":26,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":74,"samples_ran":50,"samples_unverified":24,"pointer_only_for_licence":14,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/harmonic-convolutional-networks-based-on","title":"Harmonic Convolutional Networks based on Discrete Cosine Transform","date":"2020-01-18","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":0,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/compounding-the-performance-improvements-of","title":"Compounding the Performance Improvements of Assembled Techniques in a Convolutional Neural Network","date":"2020-01-17","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":1,"samples_unverified":10,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/large-scale-learning-of-general-visual","title":"Big Transfer (BiT): General Visual Representation Learning","date":"2019-12-24","rows_on_this_dataset":2,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":10,"samples_ran":3,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/adversarial-autoaugment-1","title":"Adversarial AutoAugment","date":"2019-12-24","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/atomnas-fine-grained-end-to-end-neural-1","title":"AtomNAS: Fine-Grained End-to-End Neural Architecture Search","date":"2019-12-20","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/spinenet-learning-scale-permuted-backbone-for","title":"SpineNet: Learning Scale-Permuted Backbone for Recognition and Localization","date":"2019-12-10","rows_on_this_dataset":1,"code_links":13,"syntology":null},{"paper":"/paper/dynamic-convolution-attention-over","title":"Dynamic Convolution: Attention over Convolution Kernels","date":"2019-12-07","rows_on_this_dataset":7,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":1,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-learning-of-pretext-invariant","title":"Self-Supervised Learning of Pretext-Invariant Representations","date":"2019-12-04","rows_on_this_dataset":3,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/sgas-sequential-greedy-architecture-search","title":"SGAS: Sequential Greedy Architecture Search","date":"2019-11-30","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":1,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/whats-hidden-in-a-randomly-weighted-neural","title":"What's Hidden in a Randomly Weighted Neural Network?","date":"2019-11-29","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":2,"samples_unverified":7,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/blockwisely-supervised-neural-architecture","title":"Blockwisely Supervised Neural Architecture Search with Knowledge Distillation","date":"2019-11-29","rows_on_this_dataset":4,"code_links":1,"syntology":null},{"paper":"/paper/ghostnet-more-features-from-cheap-operations","title":"GhostNet: More Features from Cheap Operations","date":"2019-11-27","rows_on_this_dataset":5,"code_links":33,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":23,"samples_ran":6,"samples_unverified":17,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fair-darts-eliminating-unfair-advantages-in","title":"Fair DARTS: Eliminating Unfair Advantages in Differentiable Architecture Search","date":"2019-11-27","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/cspnet-a-new-backbone-that-can-enhance","title":"CSPNet: A New Backbone that can Enhance Learning Capability of CNN","date":"2019-11-27","rows_on_this_dataset":1,"code_links":123,"syntology":null},{"paper":"/paper/rigging-the-lottery-making-all-tickets-1","title":"Rigging the Lottery: Making All Tickets Winners","date":"2019-11-25","rows_on_this_dataset":4,"code_links":11,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/filter-response-normalization-layer","title":"Filter Response Normalization Layer: Eliminating Batch Dependence in the Training of Deep Neural Networks","date":"2019-11-21","rows_on_this_dataset":2,"code_links":16,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/adversarial-examples-improve-image","title":"Adversarial Examples Improve Image Recognition","date":"2019-11-21","rows_on_this_dataset":2,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/faster-autoaugment-learning-augmentation","title":"Faster AutoAugment: Learning Augmentation Strategies using Backpropagation","date":"2019-11-16","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":0,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-labelling-via-simultaneous-clustering-1","title":"Self-labelling via simultaneous clustering and representation learning","date":"2019-11-13","rows_on_this_dataset":5,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":18,"samples_ran":7,"samples_unverified":11,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/momentum-contrast-for-unsupervised-visual","title":"Momentum Contrast for Unsupervised Visual Representation Learning","date":"2019-11-13","rows_on_this_dataset":6,"code_links":44,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":42,"samples_ran":26,"samples_unverified":16,"pointer_only_for_licence":16,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-training-with-noisy-student-improves","title":"Self-training with Noisy Student improves ImageNet classification","date":"2019-11-11","rows_on_this_dataset":9,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":24,"samples_ran":5,"samples_unverified":19,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/contrastive-representation-distillation-1","title":"Contrastive Representation Distillation","date":"2019-10-23","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/on-the-adequacy-of-untuned-warmup-for","title":"On the adequacy of untuned warmup for adaptive optimization","date":"2019-10-09","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":0,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/eca-net-efficient-channel-attention-for-deep","title":"ECA-Net: Efficient Channel Attention for Deep Convolutional Neural Networks","date":"2019-10-08","rows_on_this_dataset":4,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":8,"samples_ran":1,"samples_unverified":7,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deformable-kernels-adapting-effective","title":"Deformable Kernels: Adapting Effective Receptive Fields for Object Deformation","date":"2019-10-07","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/randaugment-practical-data-augmentation-with","title":"RandAugment: Practical automated data augmentation with a reduced search space","date":"2019-09-30","rows_on_this_dataset":3,"code_links":19,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":65,"samples_ran":58,"samples_unverified":7,"pointer_only_for_licence":17,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/balanced-binary-neural-networks-with-gated","title":"Balanced Binary Neural Networks with Gated Residual","date":"2019-09-26","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/ensemble-knowledge-distillation-for-learning","title":"Ensemble Knowledge Distillation for Learning Improved and Efficient Networks","date":"2019-09-17","rows_on_this_dataset":3,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":0,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/dual-student-breaking-the-limits-of-the","title":"Dual Student: Breaking the Limits of the Teacher in Semi-supervised Learning","date":"2019-09-03","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/once-for-all-train-one-network-and-specialize","title":"Once-for-All: Train One Network and Specialize it for Efficient Deployment","date":"2019-08-26","rows_on_this_dataset":1,"code_links":10,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":34,"samples_ran":4,"samples_unverified":30,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/gated-convolutional-networks-with-hybrid","title":"Gated Convolutional Networks with Hybrid Connectivity for Image Classification","date":"2019-08-26","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mish-a-self-regularized-non-monotonic-neural","title":"Mish: A Self Regularized Non-Monotonic Activation Function","date":"2019-08-23","rows_on_this_dataset":1,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":12,"samples_ran":3,"samples_unverified":9,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/scarletnas-bridging-the-gap-between","title":"SCARLET-NAS: Bridging the Gap between Stability and Scalability in Weight-sharing Neural Architecture Search","date":"2019-08-16","rows_on_this_dataset":7,"code_links":1,"syntology":null},{"paper":"/paper/lip-local-importance-based-pooling","title":"LIP: Local Importance-based Pooling","date":"2019-08-12","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/repetitive-reprediction-deep-decipher-for","title":"Repetitive Reprediction Deep Decipher for Semi-Supervised Learning","date":"2019-08-09","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/moga-searching-beyond-mobilenetv3","title":"MoGA: Searching Beyond MobileNetV3","date":"2019-08-04","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/attentive-normalization","title":"Attentive Normalization","date":"2019-08-04","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/compact-global-descriptor-for-neural-networks","title":"Compact Global Descriptor for Neural Networks","date":"2019-07-23","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/mixnet-mixed-depthwise-convolutional-kernels","title":"MixConv: Mixed Depthwise Convolutional Kernels","date":"2019-07-22","rows_on_this_dataset":3,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":1,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pc-darts-partial-channel-connections-for","title":"PC-DARTS: Partial Channel Connections for Memory-Efficient Architecture Search","date":"2019-07-12","rows_on_this_dataset":1,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":1,"samples_unverified":10,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/large-scale-adversarial-representation","title":"Large Scale Adversarial Representation Learning","date":"2019-07-04","rows_on_this_dataset":7,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fairnas-rethinking-evaluation-fairness-of","title":"FairNAS: Rethinking Evaluation Fairness of Weight Sharing Neural Architecture Search","date":"2019-07-03","rows_on_this_dataset":6,"code_links":2,"syntology":null},{"paper":"/paper/densely-connected-search-space-for-more","title":"Densely Connected Search Space for More Flexible Neural Architecture Search","date":"2019-06-23","rows_on_this_dataset":2,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fixing-the-train-test-resolution-discrepancy","title":"Fixing the train-test resolution discrepancy","date":"2019-06-14","rows_on_this_dataset":4,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":0,"samples_unverified":2,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/contrastive-multiview-coding","title":"Contrastive Multiview Coding","date":"2019-06-13","rows_on_this_dataset":5,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":3,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/dicenet-dimension-wise-convolutions-for","title":"DiCENet: Dimension-wise Convolutions for Efficient Networks","date":"2019-06-08","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/190600910","title":"Learning Representations by Maximizing Mutual Information Across Views","date":"2019-06-03","rows_on_this_dataset":3,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":8,"samples_unverified":3,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/efficientnet-rethinking-model-scaling-for","title":"EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks","date":"2019-05-28","rows_on_this_dataset":8,"code_links":144,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":302,"samples_ran":171,"samples_unverified":131,"pointer_only_for_licence":112,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/spatial-group-wise-enhance-improving-semantic","title":"Spatial Group-wise Enhance: Improving Semantic Feature Learning in Convolutional Networks","date":"2019-05-23","rows_on_this_dataset":2,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":2,"samples_unverified":2,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/network-pruning-via-transformable","title":"Network Pruning via Transformable Architecture Search","date":"2019-05-23","rows_on_this_dataset":1,"code_links":4,"syntology":null},{"paper":"/paper/data-efficient-image-recognition-with","title":"Data-Efficient Image Recognition with Contrastive Predictive Coding","date":"2019-05-22","rows_on_this_dataset":6,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":2,"samples_ran":2,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/cutmix-regularization-strategy-to-train","title":"CutMix: Regularization Strategy to Train Strong Classifiers with Localizable Features","date":"2019-05-13","rows_on_this_dataset":2,"code_links":30,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":24,"samples_ran":17,"samples_unverified":7,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/190503670","title":"S4L: Self-Supervised Semi-Supervised Learning","date":"2019-05-09","rows_on_this_dataset":22,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":19,"samples_ran":0,"samples_unverified":19,"pointer_only_for_licence":19,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/searching-for-mobilenetv3","title":"Searching for MobileNetV3","date":"2019-05-06","rows_on_this_dataset":1,"code_links":67,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":105,"samples_ran":58,"samples_unverified":47,"pointer_only_for_licence":46,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/carafe-content-aware-reassembly-of-features","title":"CARAFE: Content-Aware ReAssembly of FEatures","date":"2019-05-06","rows_on_this_dataset":1,"code_links":3,"syntology":null},{"paper":"/paper/leveraging-large-scale-uncurated-data-for","title":"Unsupervised Pre-Training of Image Features on Non-Curated Data","date":"2019-05-03","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/billion-scale-semi-supervised-learning-for","title":"Billion-scale semi-supervised learning for image classification","date":"2019-05-02","rows_on_this_dataset":3,"code_links":4,"syntology":null},{"paper":"/paper/fast-autoaugment","title":"Fast AutoAugment","date":"2019-05-01","rows_on_this_dataset":4,"code_links":11,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":40,"samples_ran":22,"samples_unverified":18,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/self-supervised-sequence-to-sequence-asr","title":"Semi-supervised Sequence-to-sequence ASR using Unpaired Speech and Text","date":"2019-04-30","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/unsupervised-data-augmentation-1","title":"Unsupervised Data Augmentation for Consistency Training","date":"2019-04-29","rows_on_this_dataset":2,"code_links":20,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":52,"samples_ran":15,"samples_unverified":37,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/190411491","title":"Local Relation Networks for Image Recognition","date":"2019-04-25","rows_on_this_dataset":1,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/190409925","title":"Attention Augmented Convolutional Networks","date":"2019-04-22","rows_on_this_dataset":1,"code_links":14,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":3,"samples_unverified":3,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/190409460","title":"Data-Driven Neuron Allocation for Scale Aggregation Networks","date":"2019-04-20","rows_on_this_dataset":3,"code_links":1,"syntology":null},{"paper":"/paper/soft-conditional-computation","title":"CondConv: Conditionally Parameterized Convolutions for Efficient Inference","date":"2019-04-10","rows_on_this_dataset":1,"code_links":9,"syntology":null},{"paper":"/paper/drop-an-octave-reducing-spatial-redundancy-in","title":"Drop an Octave: Reducing Spatial Redundancy in Convolutional Neural Networks with Octave Convolution","date":"2019-04-10","rows_on_this_dataset":1,"code_links":28,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":34,"samples_ran":14,"samples_unverified":20,"pointer_only_for_licence":8,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/adaptively-connected-neural-networks","title":"Adaptively Connected Neural Networks","date":"2019-04-07","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/single-path-nas-designing-hardware-efficient","title":"Single-Path NAS: Designing Hardware-Efficient ConvNets in less than 4 Hours","date":"2019-04-05","rows_on_this_dataset":1,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":16,"samples_ran":1,"samples_unverified":15,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/a-comprehensive-overhaul-of-feature","title":"A Comprehensive Overhaul of Feature Distillation","date":"2019-04-03","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":1,"samples_unverified":5,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/res2net-a-new-multi-scale-backbone","title":"Res2Net: A New Multi-scale Backbone Architecture","date":"2019-04-02","rows_on_this_dataset":2,"code_links":34,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":3,"samples_unverified":6,"pointer_only_for_licence":9,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/exploring-randomly-wired-neural-networks-for","title":"Exploring Randomly Wired Neural Networks for Image Recognition","date":"2019-04-02","rows_on_this_dataset":3,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":31,"samples_ran":0,"samples_unverified":31,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/single-path-one-shot-neural-architecture","title":"Single Path One-Shot Neural Architecture Search with Uniform Sampling","date":"2019-03-31","rows_on_this_dataset":3,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":1,"samples_unverified":6,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/local-aggregation-for-unsupervised-learning","title":"Local Aggregation for Unsupervised Learning of Visual Embeddings","date":"2019-03-29","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/srm-a-style-based-recalibration-module-for","title":"SRM : A Style-based Recalibration Module for Convolutional Neural Networks","date":"2019-03-26","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/alphax-exploring-neural-architectures-with-1","title":"AlphaX: eXploring Neural Architectures with Deep Neural Networks and Monte Carlo Tree Search","date":"2019-03-26","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/sharpdarts-faster-and-more-accurate","title":"sharpDARTS: Faster and More Accurate Differentiable Architecture Search","date":"2019-03-23","rows_on_this_dataset":2,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/selective-kernel-networks","title":"Selective Kernel Networks","date":"2019-03-15","rows_on_this_dataset":1,"code_links":20,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learned-step-size-quantization","title":"Learned Step Size Quantization","date":"2019-02-21","rows_on_this_dataset":5,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":23,"samples_ran":7,"samples_unverified":16,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/multigrain-a-unified-image-embedding-for","title":"MultiGrain: a unified image embedding for classes and instances","date":"2019-02-14","rows_on_this_dataset":10,"code_links":3,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":1,"samples_unverified":3,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/graph-rise-graph-regularized-image-semantic","title":"Graph-RISE: Graph-Regularized Image Semantic Embedding","date":"2019-02-14","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/colornet-investigating-the-importance-of","title":"ColorNet: Investigating the importance of color spaces for image classification","date":"2019-02-01","rows_on_this_dataset":2,"code_links":1,"syntology":null},{"paper":"/paper/revisiting-self-supervised-visual","title":"Revisiting Self-Supervised Visual Representation Learning","date":"2019-01-25","rows_on_this_dataset":4,"code_links":6,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":0,"samples_unverified":15,"pointer_only_for_licence":7,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/feature-denoising-for-improving-adversarial","title":"Feature Denoising for Improving Adversarial Robustness","date":"2018-12-09","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/fbnet-hardware-aware-efficient-convnet-design","title":"FBNet: Hardware-Aware Efficient ConvNet Design via Differentiable Neural Architecture Search","date":"2018-12-09","rows_on_this_dataset":1,"code_links":5,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/bag-of-tricks-for-image-classification-with","title":"Bag of Tricks for Image Classification with Convolutional Neural Networks","date":"2018-12-04","rows_on_this_dataset":1,"code_links":28,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":4,"samples_unverified":11,"pointer_only_for_licence":5,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/proxylessnas-direct-neural-architecture","title":"ProxylessNAS: Direct Neural Architecture Search on Target Task and Hardware","date":"2018-12-02","rows_on_this_dataset":2,"code_links":23,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":27,"samples_ran":5,"samples_unverified":22,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/espnetv2-a-light-weight-power-efficient-and","title":"ESPNetv2: A Light-weight, Power Efficient, and General Purpose Convolutional Neural Network","date":"2018-11-28","rows_on_this_dataset":1,"code_links":10,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":2,"samples_unverified":2,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/gpipe-efficient-training-of-giant-neural","title":"GPipe: Efficient Training of Giant Neural Networks using Pipeline Parallelism","date":"2018-11-16","rows_on_this_dataset":1,"code_links":13,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":25,"samples_ran":1,"samples_unverified":24,"pointer_only_for_licence":16,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/dropblock-a-regularization-method-for","title":"DropBlock: A regularization method for convolutional networks","date":"2018-10-30","rows_on_this_dataset":1,"code_links":10,"syntology":null},{"paper":"/paper/mnasnet-platform-aware-neural-architecture","title":"MnasNet: Platform-Aware Neural Architecture Search for Mobile","date":"2018-07-31","rows_on_this_dataset":3,"code_links":29,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":2,"samples_unverified":4,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/shufflenet-v2-practical-guidelines-for","title":"ShuffleNet V2: Practical Guidelines for Efficient CNN Architecture Design","date":"2018-07-30","rows_on_this_dataset":1,"code_links":35,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":30,"samples_ran":1,"samples_unverified":29,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deep-clustering-for-unsupervised-learning-of","title":"Deep Clustering for Unsupervised Learning of Visual Features","date":"2018-07-15","rows_on_this_dataset":1,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":5,"samples_unverified":2,"pointer_only_for_licence":4,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/representation-learning-with-contrastive","title":"Representation Learning with Contrastive Predictive Coding","date":"2018-07-10","rows_on_this_dataset":3,"code_links":28,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":45,"samples_ran":29,"samples_unverified":16,"pointer_only_for_licence":22,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pcl-proposal-cluster-learning-for-weakly","title":"PCL: Proposal Cluster Learning for Weakly Supervised Object Detection","date":"2018-07-09","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":7,"samples_ran":1,"samples_unverified":6,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/an-intriguing-failing-of-convolutional-neural","title":"An Intriguing Failing of Convolutional Neural Networks and the CoordConv Solution","date":"2018-07-09","rows_on_this_dataset":1,"code_links":24,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":5,"samples_ran":4,"samples_unverified":1,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/darts-differentiable-architecture-search","title":"DARTS: Differentiable Architecture Search","date":"2018-06-24","rows_on_this_dataset":1,"code_links":59,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":156,"samples_ran":66,"samples_unverified":90,"pointer_only_for_licence":48,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unsupervised-feature-learning-via-non-1","title":"Unsupervised Feature Learning via Non-Parametric Instance Discrimination","date":"2018-06-01","rows_on_this_dataset":3,"code_links":4,"syntology":null},{"paper":"/paper/autoaugment-learning-augmentation-policies","title":"AutoAugment: Learning Augmentation Policies from Data","date":"2018-05-24","rows_on_this_dataset":2,"code_links":33,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":43,"samples_ran":6,"samples_unverified":37,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/unsupervised-feature-learning-via-non","title":"Unsupervised Feature Learning via Non-Parametric Instance-level Discrimination","date":"2018-05-05","rows_on_this_dataset":1,"code_links":15,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":23,"samples_ran":19,"samples_unverified":4,"pointer_only_for_licence":16,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/exploring-the-limits-of-weakly-supervised","title":"Exploring the Limits of Weakly Supervised Pretraining","date":"2018-05-02","rows_on_this_dataset":4,"code_links":4,"syntology":null},{"paper":"/paper/what-do-deep-networks-like-to-see","title":"What do Deep Networks Like to See?","date":"2018-03-22","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/unsupervised-representation-learning-by-1","title":"Unsupervised Representation Learning by Predicting Image Rotations","date":"2018-03-21","rows_on_this_dataset":1,"code_links":20,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":22,"samples_ran":14,"samples_unverified":8,"pointer_only_for_licence":18,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/averaging-weights-leads-to-wider-optima-and","title":"Averaging Weights Leads to Wider Optima and Better Generalization","date":"2018-03-14","rows_on_this_dataset":2,"code_links":17,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":9,"samples_ran":5,"samples_unverified":4,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/regularized-evolution-for-image-classifier","title":"Regularized Evolution for Image Classifier Architecture Search","date":"2018-02-05","rows_on_this_dataset":1,"code_links":5,"syntology":null},{"paper":"/paper/mobilenetv2-inverted-residuals-and-linear","title":"MobileNetV2: Inverted Residuals and Linear Bottlenecks","date":"2018-01-13","rows_on_this_dataset":2,"code_links":159,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":111,"samples_ran":85,"samples_unverified":26,"pointer_only_for_licence":64,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/sparse-learning-of-stochastic-dynamic","title":"Sparse learning of stochastic dynamic equations","date":"2017-12-06","rows_on_this_dataset":1,"code_links":1,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":1,"samples_unverified":2,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/progressive-neural-architecture-search","title":"Progressive Neural Architecture Search","date":"2017-12-02","rows_on_this_dataset":2,"code_links":18,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":3,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/multi-task-self-supervised-visual-learning","title":"Multi-task Self-Supervised Visual Learning","date":"2017-08-25","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/representation-learning-by-learning-to-count","title":"Representation Learning by Learning to Count","date":"2017-08-22","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/learning-transferable-architectures-for","title":"Learning Transferable Architectures for Scalable Image Recognition","date":"2017-07-21","rows_on_this_dataset":1,"code_links":17,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":0,"samples_unverified":1,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/revisiting-unreasonable-effectiveness-of-data","title":"Revisiting Unreasonable Effectiveness of Data in Deep Learning Era","date":"2017-07-10","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/shufflenet-an-extremely-efficient","title":"ShuffleNet: An Extremely Efficient Convolutional Neural Network for Mobile Devices","date":"2017-07-04","rows_on_this_dataset":1,"code_links":38,"syntology":null},{"paper":"/paper/few-example-object-detection-with-model","title":"Few-Example Object Detection with Model Communication","date":"2017-06-26","rows_on_this_dataset":1,"code_links":1,"syntology":null},{"paper":"/paper/residual-attention-network-for-image","title":"Residual Attention Network for Image Classification","date":"2017-04-23","rows_on_this_dataset":1,"code_links":19,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/mobilenets-efficient-convolutional-neural","title":"MobileNets: Efficient Convolutional Neural Networks for Mobile Vision Applications","date":"2017-04-17","rows_on_this_dataset":1,"code_links":159,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":83,"samples_ran":52,"samples_unverified":31,"pointer_only_for_licence":48,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/multiple-instance-detection-network-with","title":"Multiple Instance Detection Network with Online Instance Classifier Refinement","date":"2017-04-01","rows_on_this_dataset":1,"code_links":4,"syntology":null},{"paper":"/paper/mean-teachers-are-better-role-models-weight","title":"Mean teachers are better role models: Weight-averaged consistency targets improve semi-supervised deep learning results","date":"2017-03-06","rows_on_this_dataset":1,"code_links":8,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":6,"samples_unverified":0,"pointer_only_for_licence":6,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/paying-more-attention-to-attention-improving","title":"Paying More Attention to Attention: Improving the Performance of Convolutional Neural Networks via Attention Transfer","date":"2016-12-12","rows_on_this_dataset":2,"code_links":7,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/split-brain-autoencoders-unsupervised","title":"Split-Brain Autoencoders: Unsupervised Learning by Cross-Channel Prediction","date":"2016-11-29","rows_on_this_dataset":1,"code_links":2,"syntology":null},{"paper":"/paper/weakly-supervised-cascaded-convolutional","title":"Weakly Supervised Cascaded Convolutional Networks","date":"2016-11-24","rows_on_this_dataset":1,"code_links":0,"syntology":null},{"paper":"/paper/aggregated-residual-transformations-for-deep","title":"Aggregated Residual Transformations for Deep Neural Networks","date":"2016-11-16","rows_on_this_dataset":1,"code_links":61,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":80,"samples_ran":34,"samples_unverified":46,"pointer_only_for_licence":13,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/xception-deep-learning-with-depthwise","title":"Xception: Deep Learning with Depthwise Separable Convolutions","date":"2016-10-07","rows_on_this_dataset":1,"code_links":41,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":15,"samples_ran":1,"samples_unverified":14,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/pruning-filters-for-efficient-convnets","title":"Pruning Filters for Efficient ConvNets","date":"2016-08-31","rows_on_this_dataset":3,"code_links":21,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":36,"samples_ran":20,"samples_unverified":16,"pointer_only_for_licence":20,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/densely-connected-convolutional-networks","title":"Densely Connected Convolutional Networks","date":"2016-08-25","rows_on_this_dataset":4,"code_links":146,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":71,"samples_ran":18,"samples_unverified":53,"pointer_only_for_licence":7,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/lets-keep-it-simple-using-simple","title":"Lets keep it simple, Using simple architectures to outperform deeper and more complex architectures","date":"2016-08-22","rows_on_this_dataset":8,"code_links":9,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":11,"samples_ran":3,"samples_unverified":8,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/fractalnet-ultra-deep-neural-networks-without","title":"FractalNet: Ultra-Deep Neural Networks without Residuals","date":"2016-05-24","rows_on_this_dataset":1,"code_links":4,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":6,"samples_ran":4,"samples_unverified":2,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/wide-residual-networks","title":"Wide Residual Networks","date":"2016-05-23","rows_on_this_dataset":1,"code_links":72,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":96,"samples_ran":60,"samples_unverified":36,"pointer_only_for_licence":46,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/colorful-image-colorization","title":"Colorful Image Colorization","date":"2016-03-28","rows_on_this_dataset":1,"code_links":39,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":73,"samples_ran":32,"samples_unverified":41,"pointer_only_for_licence":41,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/identity-mappings-in-deep-residual-networks","title":"Identity Mappings in Deep Residual Networks","date":"2016-03-16","rows_on_this_dataset":1,"code_links":54,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":25,"samples_ran":3,"samples_unverified":22,"pointer_only_for_licence":1,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/synthesized-classifiers-for-zero-shot","title":"Synthesized Classifiers for Zero-Shot Learning","date":"2016-03-02","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/squeezenet-alexnet-level-accuracy-with-50x","title":"SqueezeNet: AlexNet-level accuracy with 50x fewer parameters and <0.5MB model size","date":"2016-02-24","rows_on_this_dataset":1,"code_links":59,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":4,"samples_ran":4,"samples_unverified":0,"pointer_only_for_licence":2,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/inception-v4-inception-resnet-and-the-impact","title":"Inception-v4, Inception-ResNet and the Impact of Residual Connections on Learning","date":"2016-02-23","rows_on_this_dataset":1,"code_links":87,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":1,"samples_ran":1,"samples_unverified":0,"pointer_only_for_licence":0,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/deep-residual-learning-for-image-recognition","title":"Deep Residual Learning for Image Recognition","date":"2015-12-10","rows_on_this_dataset":3,"code_links":484,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":377,"samples_ran":230,"samples_unverified":147,"pointer_only_for_licence":187,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/firecaffe-near-linear-acceleration-of-deep","title":"FireCaffe: near-linear acceleration of deep neural network training on compute clusters","date":"2015-10-31","rows_on_this_dataset":2,"code_links":0,"syntology":null},{"paper":"/paper/distilling-the-knowledge-in-a-neural-network","title":"Distilling the Knowledge in a Neural Network","date":"2015-03-09","rows_on_this_dataset":2,"code_links":64,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":37,"samples_ran":15,"samples_unverified":22,"pointer_only_for_licence":10,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}},{"paper":"/paper/zero-shot-learning-by-convex-combination-of","title":"Zero-Shot Learning by Convex Combination of Semantic Embeddings","date":"2013-12-19","rows_on_this_dataset":1,"code_links":2,"syntology":{"read_at":"2026-09-24T18:15:14+00:00","samples_harvested":3,"samples_ran":0,"samples_unverified":3,"pointer_only_for_licence":3,"claim":"Per-sample execution on synthesized fixtures; not a correctness claim."}}],"syntology_totals":{"read_at":"2026-09-24T18:15:14+00:00","papers_with_samples":382,"samples_harvested":6586,"samples_ran":3492,"samples_unverified":3094,"pointer_only_for_licence":1942,"papers_with_no_sample_that_ran":51,"note":"the per-paper counts above, summed; not a rate"},"papers_note":"The archive never published its papers-using-dataset list; these are papers with a leaderboard row on this dataset's benchmarks."}