Papers › Improving Augmentation and Evaluation Schemes for Semantic Image Synthesis

Improving Augmentation and Evaluation Schemes for Semantic Image Synthesis

25 Nov 2020arXiv:2011.12636archive 2025-07-28

Prateek Katiyar, Anna Khoreva

Despite data augmentation being a de facto technique for boosting the performance of deep neural networks, little attention has been paid to developing augmentation strategies for generative adversarial networks (GANs). To this end, we introduce a novel augmentation scheme designed specifically for GAN-based semantic image synthesis models. We propose to randomly warp object shapes in the semantic label maps used as an input to the generator. The local shape discrepancies between the warped and non-warped label maps and images enable the GAN to learn better the structural and geometric details of the scene and thus to improve the quality of generated images. While benchmarking the augmented GAN models against their vanilla counterparts, we discover that the quantification metrics reported in the previous semantic image synthesis studies are strongly biased towards specific semantic classes as they are derived via an external pre-trained segmentation network. We therefore propose to improve the established semantic image synthesis evaluation scheme by analyzing separately the performance of generated images on the biased and unbiased classes for the given segmentation network. Finally, we show strong quantitative and qualitative improvements obtained with our augmentation scheme, on both class splits, using state-of-the-art semantic image synthesis models across three different datasets. On average across COCO-Stuff, ADE20K and Cityscapes datasets, the augmented models outperform their vanilla counterparts by ~3 mIoU and ~10 FID points.

PaperPDF

Code

No code repository is listed for this paper in the archive or in Syntology's graph.

Code Syntology ran Syntology

Not run by Syntology. Nothing on this page verifies that the listed code works.

Tasks

BenchmarkingData AugmentationImage GenerationImage-to-Image Translation

Results from the paper archive 2025-07-28

TaskDatasetModelMetricValueRank at snapshotLeaderboardReport
Image-to-Image Translation ADE20K Labels-to-Photos CC-FPSE-AUG Accuracy 83% #6 of 16 Archive leaderboard report
Image-to-Image Translation ADE20K Labels-to-Photos CC-FPSE-AUG FID 32.6 #6 of 16 Archive leaderboard report
Image-to-Image Translation ADE20K Labels-to-Photos CC-FPSE-AUG mIoU 44 #6 of 16 Archive leaderboard report
Image-to-Image Translation ADE20K Labels-to-Photos Pix2PixHD-AUG Accuracy 77.9% #13 of 16 Archive leaderboard report
Image-to-Image Translation ADE20K Labels-to-Photos Pix2PixHD-AUG FID 41.5 #13 of 16 Archive leaderboard report
Image-to-Image Translation COCO-Stuff Labels-to-Photos CC-FPSE-AUG Accuracy 71.5 #7 of 15 Archive leaderboard report
Image-to-Image Translation COCO-Stuff Labels-to-Photos CC-FPSE-AUG FID 19.1 #7 of 15 Archive leaderboard report
Image-to-Image Translation COCO-Stuff Labels-to-Photos CC-FPSE-AUG mIoU 42.1 #7 of 15 Archive leaderboard report
Image-to-Image Translation COCO-Stuff Labels-to-Photos Pix2PixHD-AUG Accuracy 54.1 #13 of 15 Archive leaderboard report
Image-to-Image Translation COCO-Stuff Labels-to-Photos Pix2PixHD-AUG FID 54.2 #13 of 15 Archive leaderboard report
Image-to-Image Translation COCO-Stuff Labels-to-Photos Pix2PixHD-AUG mIoU 21.9 #13 of 15 Archive leaderboard report
Image-to-Image Translation Cityscapes Labels-to-Photo CC-FPSE-AUG Accuracy 93.5 #7 of 21 Archive leaderboard report
Image-to-Image Translation Cityscapes Labels-to-Photo CC-FPSE-AUG FID 52.1 #7 of 21 Archive leaderboard report
Image-to-Image Translation Cityscapes Labels-to-Photo CC-FPSE-AUG mIoU 63.1 #7 of 21 Archive leaderboard report
Image-to-Image Translation Cityscapes Labels-to-Photo Pix2PixHD-AUG Accuracy 92.7 #10 of 21 Archive leaderboard report
Image-to-Image Translation Cityscapes Labels-to-Photo Pix2PixHD-AUG FID 72.7 #10 of 21 Archive leaderboard report
Image-to-Image Translation Cityscapes Labels-to-Photo Pix2PixHD-AUG mIoU 58 #10 of 21 Archive leaderboard report

Ranks are positions in the archive's leaderboards as they stood at the 2025-07-28 snapshot. Results published since then are not among these rows, so a rank here is not a current standing.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections