TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Image-to-Image Translation	ADE20K Labels-to-Photos	USIS	mIoU	17.38	# 10
Image-to-Image Translation	ADE20K Labels-to-Photos	USIS	FID	33.2	# 8
Image-to-Image Translation	Cityscapes Labels-to-Photo	USIS	mIoU	44.78	# 13
Image-to-Image Translation	Cityscapes Labels-to-Photo	USIS	FID	53.67	# 7
Image-to-Image Translation	COCO-Stuff Labels-to-Photos	USIS	mIoU	14.06	# 8
Image-to-Image Translation	COCO-Stuff Labels-to-Photos	USIS	FID	27.8	# 10

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/usis-unsupervised-semantic-image-synthesis/image-to-image-translation-on-ade20k-labels)](https://paperswithcode.com/sota/image-to-image-translation-on-ade20k-labels?p=usis-unsupervised-semantic-image-synthesis)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/usis-unsupervised-semantic-image-synthesis/image-to-image-translation-on-coco-stuff)](https://paperswithcode.com/sota/image-to-image-translation-on-coco-stuff?p=usis-unsupervised-semantic-image-synthesis)`
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/usis-unsupervised-semantic-image-synthesis/image-to-image-translation-on-cityscapes)](https://paperswithcode.com/sota/image-to-image-translation-on-cityscapes?p=usis-unsupervised-semantic-image-synthesis)`

USIS: Unsupervised Semantic Image Synthesis

29 Sep 2021 · George Eskandar, Mohamed Abdelsamad, Karim Armanious, Bin Yang ·

Semantic Image Synthesis (SIS) is a subclass of image-to-image translation where a photorealistic image is synthesized from a segmentation mask. SIS has mostly been addressed as a supervised problem. However, state-of-the-art methods depend on a huge amount of labeled data and cannot be applied in an unpaired setting. On the other hand, generic unpaired image-to-image translation frameworks underperform in comparison, because they color-code semantic layouts and feed them to traditional convolutional networks, which then learn correspondences in appearance instead of semantic content. In this initial work, we propose a new Unsupervised paradigm for Semantic Image Synthesis (USIS) as a first step towards closing the performance gap between paired and unpaired settings. Notably, the framework deploys a SPADE generator that learns to output images with visually separable semantic classes using a self-supervised segmentation loss. Furthermore, in order to match the color and texture distribution of real images without losing high-frequency information, we propose to use whole image wavelet-based discrimination. We test our methodology on 3 challenging datasets and demonstrate its ability to generate multimodal photorealistic images with an improved quality in the unpaired setting.

PDF Abstract

Code

Add Remove Mark official

GeorgeEskandar/USIS-Unsupervised-Se… official

Tasks

Add Remove

Image Generation

Image-to-Image Translation

Translation

Datasets

Cityscapes

ADE20K

COCO-Stuff

Results from the Paper

Edit

Ranked #10 on Image-to-Image Translation on COCO-Stuff Labels-to-Photos

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Image-to-Image Translation	ADE20K Labels-to-Photos	USIS	mIoU	17.38	# 10	Compare
Image-to-Image Translation	ADE20K Labels-to-Photos	USIS	FID	33.2	# 8	Compare
Image-to-Image Translation	Cityscapes Labels-to-Photo	USIS	mIoU	44.78	# 13	Compare
Image-to-Image Translation	Cityscapes Labels-to-Photo	USIS	FID	53.67	# 7	Compare
Image-to-Image Translation	COCO-Stuff Labels-to-Photos	USIS	mIoU	14.06	# 8	Compare
Image-to-Image Translation	COCO-Stuff Labels-to-Photos	USIS	FID	27.8	# 10	Compare

Methods

Add Remove

SPADE • Test

Edit Social Preview

USIS: Unsupervised Semantic Image Synthesis

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove