{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/learning-from-extrinsic-and-intrinsic","title":"Learning from Extrinsic and Intrinsic Supervisions for Domain Generalization","arxiv_id":"2007.09316","date":"2020-07-18","proceeding":"ECCV 2020 8","authors":["Shujun Wang","Lequan Yu","Caizi Li","Chi-Wing Fu","Pheng-Ann Heng"],"abstract":"The generalization capability of neural networks across domains is crucial for real-world applications. We argue that a generalized object recognition system should well understand the relationships among different images and also the images themselves at the same time. To this end, we present a new domain generalization framework that learns how to generalize across domains simultaneously from extrinsic relationship supervision and intrinsic self-supervision for images from multi-source domains. To be specific, we formulate our framework with feature embedding using a multi-task learning paradigm. Besides conducting the common supervised recognition task, we seamlessly integrate a momentum metric learning task and a self-supervised auxiliary task to collectively utilize the extrinsic supervision and intrinsic supervision. Also, we develop an effective momentum metric learning scheme with K-hard negative mining to boost the network to capture image relationship for domain generalization. We demonstrate the effectiveness of our approach on two standard object recognition benchmarks VLCS and PACS, and show that our methods achieve state-of-the-art performance.","url_abs":"https://arxiv.org/abs/2007.09316v1","url_pdf":"https://arxiv.org/pdf/2007.09316v1.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[],"tasks":[{"task_slug":"anomaly-detection","task_name":"Anomaly Detection"},{"task_slug":"domain-generalization","task_name":"Domain Generalization"},{"task_slug":"metric-learning","task_name":"Metric Learning"},{"task_slug":"multi-task-learning","task_name":"Multi-Task Learning"},{"task_slug":"object-recognition","task_name":"Object Recognition"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/anomaly-detection-on-mvtec-ad-textures-domain","task":"Anomaly Detection","dataset":"MVTec AD Textures Domain Generalization","model":"EISNet+","rank_in_archive_order":3,"of":3,"metrics":{"Detection AUROC":"90.9"},"uses_additional_data":false},{"leaderboard":"/sota/domain-generalization-on-pacs-2","task":"Domain Generalization","dataset":"PACS","model":"EISNet (Resnet-50)","rank_in_archive_order":49,"of":133,"metrics":{"Average Accuracy":"85.84"},"uses_additional_data":false},{"leaderboard":"/sota/domain-generalization-on-pacs-2","task":"Domain Generalization","dataset":"PACS","model":"EISNet (Resnet-18)","rank_in_archive_order":80,"of":133,"metrics":{"Average Accuracy":"82.15"},"uses_additional_data":false}],"syntology":{"syntology_url":"https://syntology.ai/paper/2007.09316","atlas_url":"https://app.syntology.ai/?focus=2007.09316","mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}