{"url":"/task/mitigating-contextual-bias","name":"Mitigating Contextual Bias","slug":"mitigating-contextual-bias","description_markdown":"The contextual bias split of the Aircraft dataset was constructed by Dunlap et al. [https://arxiv.org/abs/2305.16289].\r\nThe split uses two visually similar classes: Boeing-767 and Airbus-322. Each image in this split is categorized as “sky”, “grass”, or “road” depending on its background, with ambiguous examples filtered out. The bias in the dataset is introduced by training on 400 samples where Airbus aircraft are exclusively associated with road backgrounds and Boeing aircraft with grass backgrounds, although both types may appear against sky backgrounds.","categories":[{"name":"Adversarial","url":"/area/adversarial"},{"name":"Audio","url":"/area/audio"},{"name":"Computer Code","url":"/area/computer-code"},{"name":"Computer Vision","url":"/area/computer-vision"},{"name":"Graphs","url":"/area/graphs"},{"name":"Medical","url":"/area/medical"},{"name":"Methodology","url":"/area/methodology"},{"name":"Miscellaneous","url":"/area/miscellaneous"},{"name":"Natural Language Processing","url":"/area/natural-language-processing"},{"name":"Reasoning","url":"/area/reasoning"},{"name":"Time Series","url":"/area/time-series"}],"source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","slug_source":"archive_url"},"counts":{"papers_tagged":4,"papers_with_code":4,"benchmarks":1,"benchmark_tables_in_archive":1,"benchmark_tables_shown":1,"benchmark_tables_withheld_as_spam":0,"benchmark_definition":"a leaderboard table with at least one row; benchmark_tables_shown also counts the zero-row tables; benchmark_tables_in_archive adds the tables withheld as spam","datasets":1,"subtasks":0,"parent_tasks":1},"benchmarks":[{"leaderboard":"/sota/mitigating-contextual-bias-on-fgvc-aircraft","slug":"mitigating-contextual-bias-on-fgvc-aircraft","dataset":"FGVC Aircraft","dataset_url":"/dataset/fgvc-aircraft-1","rows_in_archive":4,"metrics":["Top-1 Accuracy (%)","OOD Accuracy (%)"],"first_row_in_archive_order":{"model":"CAL + SaSPA","paper_title":"Advancing Fine-Grained Classification by Structure and Subject Preserving Augmentation","paper_url":"/paper/advancing-fine-grained-classification-by","paper_date":"2024-06-20","arxiv_id":"2406.14551","code_links":[{"title":"eyalmichaeli/saspa-aug","url":"https://github.com/eyalmichaeli/saspa-aug"}],"syntology":{"n":1,"n_ran":0,"n_unverified":1,"n_pointer_only":0}}}],"datasets":[{"url":"/dataset/fgvc-aircraft-1","name":"FGVC-Aircraft","full_name":"","num_papers_in_archive":520}],"subtasks":[],"parent_tasks":[{"url":"/task/classification-1","name":"Classification"}],"papers":{"order":"repositories listed in the archive (desc), then date (desc); the archive holds no stars","population":"papers tagged with this task that list at least one repository in the archive","shown":4,"of":4,"tagged_in_all":4,"items":[{"url":"/paper/advancing-fine-grained-classification-by","title":"Advancing Fine-Grained Classification by Structure and Subject Preserving Augmentation","date":"2024-06-20","arxiv_id":"2406.14551","repositories_listed":1,"syntology":{"n":1,"n_ran":0,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/denetdm-debiasing-by-network-depth-modulation","title":"DeNetDM: Debiasing by Network Depth Modulation","date":"2024-03-28","arxiv_id":"2403.19863","repositories_listed":1,"syntology":{"n":7,"n_ran":6,"n_unverified":1,"n_pointer_only":0}},{"url":"/paper/is-synthetic-data-from-diffusion-models-ready","title":"Is Synthetic Data From Diffusion Models Ready for Knowledge Distillation?","date":"2023-05-22","arxiv_id":"2305.12954","repositories_listed":1,"syntology":{"n":1,"n_ran":1,"n_unverified":0,"n_pointer_only":1}},{"url":"/paper/counterfactual-attention-learning-for-fine","title":"Counterfactual Attention Learning for Fine-Grained Visual Categorization and Re-identification","date":"2021-08-19","arxiv_id":"2108.08728","repositories_listed":1,"syntology":{"n":11,"n_ran":8,"n_unverified":3,"n_pointer_only":0}}],"syntology_records":4,"syntology_note":"a paper without a record is not a recorded non-run: it may lack an arXiv id or simply be absent from the graph layer"},"description_links":{"kept":0,"unwrapped_to_text":0,"bare_urls_linked":0,"relative_images_dropped":0,"rule":"internal links are kept only when the target slug exists in the catalog"},"syntology":{"read_at":"2026-09-25T09:33:49+00:00","claim":"Per-sample execution status on synthesized fixtures ('ran N of M samples'); not a correctness claim and not a ranking signal.","status_vocabulary":{"ran_honours":"ran, honoured the contract we drafted","ran_violates":"ran, violated the contract we drafted","ran_draft_wrong":"ran; our contract draft was wrong, not the code","ran_fixture":"ran; our fixture could not drive it","ran":"ran on a synthesized input","unverified":"unverified (harvested, no recorded run)"}},"not_shown":{"libraries":"the archive has no per-task library table","trend_sparklines":"the Trend column of the benchmarks table was a rendered image; it is not in the archive","social_and_latest_sorts":"stars and social signals are not in the archive"}}