{"about":{"site":"https://codewithpapers.app","non_affiliation":"Code with Papers and Syntology are not affiliated with, endorsed by, or sponsored by Papers with Code, Meta, or the pwc-archive mirror.","licence":"CC BY-SA 4.0","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","attribution":"https://codewithpapers.app/attribution","modified":"archive material modified by Syntology; see the attribution page"},"url":"/paper/procst-boosting-semantic-segmentation-using","title":"ProCST: Boosting Semantic Segmentation Using Progressive Cyclic Style-Transfer","arxiv_id":"2204.11891","date":"2022-04-25","proceeding":null,"authors":["Shahaf Ettedgui","Shady Abu-Hussein","Raja Giryes"],"abstract":"Using synthetic data for training neural networks that achieve good performance on real-world data is an important task as it can reduce the need for costly data annotation. Yet, synthetic and real world data have a domain gap. Reducing this gap, also known as domain adaptation, has been widely studied in recent years. Closing the domain gap between the source (synthetic) and target (real) data by directly performing the adaptation between the two is challenging. In this work, we propose a novel two-stage framework for improving domain adaptation techniques on image data. In the first stage, we progressively train a multi-scale neural network to perform image translation from the source domain to the target domain. We denote the new transformed data as \"Source in Target\" (SiT). Then, we insert the generated SiT data as the input to any standard UDA approach. This new data has a reduced domain gap from the desired target domain, which facilitates the applied UDA approach to close the gap further. We emphasize the effectiveness of our method via a comparison to other leading UDA and image-to-image translation techniques when used as SiT generators. Moreover, we demonstrate the improvement of our framework with three state-of-the-art UDA methods for semantic segmentation, HRDA, DAFormer and ProDA, on two UDA tasks, GTA5 to Cityscapes and Synthia to Cityscapes.","url_abs":"https://arxiv.org/abs/2204.11891v2","url_pdf":"https://arxiv.org/pdf/2204.11891v2.pdf","source":{"archive":"pwc-archive (Hugging Face), CC BY-SA 4.0","snapshot":"2025-07-28","licence_url":"https://creativecommons.org/licenses/by-sa/4.0/legalcode","row_kind":"abstracts"},"code_links":[{"paper_slug":"procst-boosting-semantic-segmentation-using","repo_url":"https://github.com/shahaf1313/procst","is_official":1,"mentioned_in_paper":1,"mentioned_in_github":1,"framework":"pytorch","reach":null}],"tasks":[{"task_slug":"domain-adaptation","task_name":"Domain Adaptation"},{"task_slug":"image-to-image-translation","task_name":"Image-to-Image Translation"},{"task_slug":"semantic-segmentation","task_name":"Semantic Segmentation"},{"task_slug":"style-transfer","task_name":"Style Transfer"},{"task_slug":"synthetic-to-real-translation","task_name":"Synthetic-to-Real Translation"},{"task_slug":"translation","task_name":"Translation"},{"task_slug":"unsupervised-domain-adaptation","task_name":"Unsupervised Domain Adaptation"}],"methods":[],"datasets_introduced":[],"methods_introduced":[],"results":[{"leaderboard":"/sota/domain-adaptation-on-gta5-to-cityscapes","task":"Domain Adaptation","dataset":"GTA5 to Cityscapes","model":"DAFormer + ProCST","rank_in_archive_order":9,"of":28,"metrics":{"mIoU":"69.4"},"uses_additional_data":true},{"leaderboard":"/sota/domain-adaptation-on-synthia-to-cityscapes","task":"Domain Adaptation","dataset":"SYNTHIA-to-Cityscapes","model":"DAFormer + ProCST","rank_in_archive_order":10,"of":33,"metrics":{"mIoU":"61.6"},"uses_additional_data":true},{"leaderboard":"/sota/image-to-image-translation-on-gtav-to","task":"Image-to-Image Translation","dataset":"GTAV-to-Cityscapes Labels","model":"DAFormer + ProCST","rank_in_archive_order":7,"of":22,"metrics":{"mIoU":"69.4"},"uses_additional_data":true},{"leaderboard":"/sota/image-to-image-translation-on-synthia-to","task":"Image-to-Image Translation","dataset":"SYNTHIA-to-Cityscapes","model":"DAFormer + ProCST","rank_in_archive_order":6,"of":28,"metrics":{"mIoU (13 classes)":"68.2"},"uses_additional_data":true},{"leaderboard":"/sota/semantic-segmentation-on-gtav-to-cityscapes-1","task":"Semantic Segmentation","dataset":"GTAV-to-Cityscapes Labels","model":"DAFormer + ProCST","rank_in_archive_order":5,"of":12,"metrics":{"mIoU":"69.4"},"uses_additional_data":true},{"leaderboard":"/sota/semantic-segmentation-on-synthia-to","task":"Semantic Segmentation","dataset":"SYNTHIA-to-Cityscapes","model":"DAFormer + ProCST","rank_in_archive_order":5,"of":7,"metrics":{"Mean IoU":"61.6"},"uses_additional_data":true},{"leaderboard":"/sota/synthetic-to-real-translation-on-gtav-to","task":"Synthetic-to-Real Translation","dataset":"GTAV-to-Cityscapes Labels","model":"DAFormer + ProCST","rank_in_archive_order":10,"of":73,"metrics":{"mIoU":"69.4"},"uses_additional_data":true},{"leaderboard":"/sota/synthetic-to-real-translation-on-synthia-to-1","task":"Synthetic-to-Real Translation","dataset":"SYNTHIA-to-Cityscapes","model":"DAFormer + ProCST","rank_in_archive_order":7,"of":38,"metrics":{"MIoU (16 classes)":"61.6"},"uses_additional_data":true},{"leaderboard":"/sota/unsupervised-domain-adaptation-on-gtav-to","task":"Unsupervised Domain Adaptation","dataset":"GTAV-to-Cityscapes Labels","model":"DAFormer + ProCST","rank_in_archive_order":9,"of":20,"metrics":{"mIoU":"69.4"},"uses_additional_data":true},{"leaderboard":"/sota/unsupervised-domain-adaptation-on-synthia-to","task":"Unsupervised Domain Adaptation","dataset":"SYNTHIA-to-Cityscapes","model":"DAFormer + ProCST","rank_in_archive_order":8,"of":23,"metrics":{"mIoU (13 classes)":"68.2"},"uses_additional_data":true}],"syntology":{"atlas_url":null,"mcp":null,"developers":"https://syntology.ai/developers"},"arxiv_metadata":null,"syntology_extracted_results":null}