{"url":"/sota/image-classification-on-dtd","task":{"name":"Image Classification","url":"/task/image-classification","note":null},"dataset":{"name":"DTD","url":"/dataset/dtd"},"category":"Computer Vision","categories":["Adversarial","Computer Vision"],"category_note":null,"description":"**Image Classification** is a fundamental task in vision recognition that aims to understand and categorize an image as a whole under a specific label. Unlike [object detection](/task/object-detection), which involves classification and location of multiple objects within an image, image classification typically pertains to single-object images. When the classification becomes highly detailed or reaches instance-level, it is often referred to as [image retrieval](/task/image-retrieval), which also involves finding similar images in a large database.\r\n\r\n\r\n<span class=\"description-source\">Source: [Metamorphic Testing for Object Detection Systems ](https://arxiv.org/abs/1912.12162)</span>","description_from":"task","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","rank":"the archive's row order at snapshot; not re-ranked","rows_end_at":"2025-07-28","rows_withheld_as_spam":0,"metric_values":"the archive's strings, untouched"},"metrics":["Accuracy"],"metric_direction":{"note":"inferred from the metric name only (the archive records no direction); null = not inferred, chart draws points only","by_metric":{"Accuracy":"higher"}},"counts":{"rows":11,"rows_with_code":11,"rows_with_paper_page":11,"rows_dated":11,"rows_using_additional_data":10},"rows":[{"rank_in_archive_order":1,"model":"Linear FT(ViT-L/14)","metrics":{"Accuracy":"90.0"},"uses_additional_data":true,"paper_date":"2023-05-22","paper":"/paper/task-arithmetic-in-the-tangent-space-improved","paper_url":"https://arxiv.org/abs/2305.12827v3","paper_title":"Task Arithmetic in the Tangent Space: Improved Editing of Pre-Trained Models","code":"https://github.com/gortizji/tangent_task_arithmetic","n_code_links":1,"syntology":{"n_ran":1,"n_unverified":3,"n_samples":4,"n_pointer_only_licence":0}},{"rank_in_archive_order":2,"model":"RADAM (ConvNeXt-L)","metrics":{"Accuracy":"84.0"},"uses_additional_data":true,"paper_date":"2023-03-08","paper":"/paper/radam-texture-recognition-through-randomized-1","paper_url":"https://arxiv.org/abs/2303.04554v1","paper_title":"RADAM: Texture Recognition through Randomized Aggregated Encoding of Deep Activation Maps","code":"https://github.com/scabini/RADAM","n_code_links":1,"syntology":null},{"rank_in_archive_order":3,"model":"µ2Net+ (ViT-L/16)","metrics":{"Accuracy":"82.23"},"uses_additional_data":true,"paper_date":"2022-09-15","paper":"/paper/a-continual-development-methodology-for-large","paper_url":"https://arxiv.org/abs/2209.07326v3","paper_title":"A Continual Development Methodology for Large-scale Multitask Dynamic ML Systems","code":"https://github.com/google-research/google-research/tree/master/muNet","n_code_links":1,"syntology":null},{"rank_in_archive_order":4,"model":"Bamboo (ViT-B/16)","metrics":{"Accuracy":"81.9"},"uses_additional_data":true,"paper_date":"2022-03-15","paper":"/paper/bamboo-building-mega-scale-vision-dataset","paper_url":"https://arxiv.org/abs/2203.07845v2","paper_title":"Bamboo: Building Mega-Scale Vision Dataset Continually with Human-Machine Synergy","code":"https://github.com/zhangyuanhan-ai/bamboo","n_code_links":2,"syntology":{"n_ran":3,"n_unverified":0,"n_samples":3,"n_pointer_only_licence":3}},{"rank_in_archive_order":5,"model":"µ2Net (ViT-L/16)","metrics":{"Accuracy":"81.0"},"uses_additional_data":true,"paper_date":"2022-05-25","paper":"/paper/an-evolutionary-approach-to-dynamic","paper_url":"https://arxiv.org/abs/2205.12755v6","paper_title":"An Evolutionary Approach to Dynamic Introduction of Tasks in Large-scale Multitask Learning Systems","code":"https://github.com/google-research/google-research/tree/master/muNet","n_code_links":1,"syntology":null},{"rank_in_archive_order":6,"model":"SEER (RegNet10B - linear eval)","metrics":{"Accuracy":"80.5"},"uses_additional_data":true,"paper_date":"2022-02-16","paper":"/paper/vision-models-are-more-robust-and-fair-when","paper_url":"https://arxiv.org/abs/2202.08360v2","paper_title":"Vision Models Are More Robust And Fair When Pretrained On Uncurated Images Without Supervision","code":"https://github.com/facebookresearch/vissl","n_code_links":1,"syntology":null},{"rank_in_archive_order":7,"model":"Inceptionv4","metrics":{"Accuracy":"79.79"},"uses_additional_data":true,"paper_date":"2021-07-19","paper":"/paper/non-binary-deep-transfer-learning-for","paper_url":"https://arxiv.org/abs/2107.08585v2","paper_title":"Non-binary deep transfer learning for image classification","code":"https://github.com/XuyangSHEN/Non-binary-deep-transfer-learning-for-image-classification","n_code_links":1,"syntology":null},{"rank_in_archive_order":8,"model":"TWIST (ResNet-50)","metrics":{"Accuracy":"76.6"},"uses_additional_data":true,"paper_date":"2021-10-14","paper":"/paper/self-supervised-learning-by-estimating-twin-1","paper_url":"https://arxiv.org/abs/2110.07402v4","paper_title":"Self-Supervised Learning by Estimating Twin Class Distributions","code":"https://github.com/bytedance/TWIST","n_code_links":2,"syntology":{"n_ran":5,"n_unverified":10,"n_samples":15,"n_pointer_only_licence":0}},{"rank_in_archive_order":9,"model":"TransBoost-ResNet50","metrics":{"Accuracy":"76.49"},"uses_additional_data":true,"paper_date":"2022-05-26","paper":"/paper/transboost-improving-the-best-imagenet","paper_url":"https://arxiv.org/abs/2205.13331v4","paper_title":"TransBoost: Improving the Best ImageNet Performance using Deep Transduction","code":"https://github.com/omerb01/transboost","n_code_links":1,"syntology":null},{"rank_in_archive_order":10,"model":"NNCLR","metrics":{"Accuracy":"75.5"},"uses_additional_data":true,"paper_date":"2021-04-29","paper":"/paper/with-a-little-help-from-my-friends-nearest","paper_url":"https://arxiv.org/abs/2104.14548v2","paper_title":"With a Little Help from My Friends: Nearest-Neighbor Contrastive Learning of Visual Representations","code":"https://github.com/lightly-ai/lightly","n_code_links":4,"syntology":{"n_ran":4,"n_unverified":1,"n_samples":5,"n_pointer_only_licence":0}},{"rank_in_archive_order":11,"model":"Inceptionv4 (random initialization)","metrics":{"Accuracy":"66.8"},"uses_additional_data":false,"paper_date":"2021-07-19","paper":"/paper/non-binary-deep-transfer-learning-for","paper_url":"https://arxiv.org/abs/2107.08585v2","paper_title":"Non-binary deep transfer learning for image classification","code":"https://github.com/XuyangSHEN/Non-binary-deep-transfer-learning-for-image-classification","n_code_links":1,"syntology":null}],"since_archive":{"present":false,"note":"No Syntology-extracted rows are published in this build."},"syntology":{"read_at":"2026-09-24T18:15:14+00:00","claim":"Per row: N of M harvested code samples from that row's paper executed on a synthesized fixture; the other M-N are unverified. Not a reproduction of the row's number; not a correctness claim. n_pointer_only_licence counts samples the site points at rather than redistributes (a licence axis, independent of ran/unverified).","rows_with_graph_line":4,"rows_with_any_sample_ran":4,"distinct_papers_with_graph_line":4,"distinct_papers_with_any_sample_ran":4,"samples_over_distinct_papers":{"n_ran":13,"n_unverified":14,"n_samples":27,"n_pointer_only_licence":3,"note":"each paper (arXiv id) counted once, however many rows it is behind; this is the page-level figure"},"samples_row_weighted":{"n_ran":13,"n_unverified":14,"n_samples":27,"n_pointer_only_licence":3,"note":"row-weighted: a paper behind several rows is counted once per row; inflated relative to samples_over_distinct_papers by design, kept for readers summing the per-row syntology blocks"}}}