TASK	DATASET	MODEL	METRIC NAME	METRIC VALUE	GLOBAL RANK
Text-to-Image Generation	MS COCO	StackGAN + OP	FID	55.30	# 66
Text-to-Image Generation	MS COCO	StackGAN + OP	Inception score	12.12	# 22
Text-to-Image Generation	MS COCO	AttnGAN + OP	FID	33.35	# 62
Text-to-Image Generation	MS COCO	AttnGAN + OP	Inception score	24.76	# 16
Text-to-Image Generation	MS COCO	AttnGAN + OP	SOA-C	25.46	# 4

Badge	Markdown
	`[![PWC](https://img.shields.io/endpoint.svg?url=https://paperswithcode.com/badge/generating-multiple-objects-at-spatially/text-to-image-generation-on-coco)](https://paperswithcode.com/sota/text-to-image-generation-on-coco?p=generating-multiple-objects-at-spatially)`

Generating Multiple Objects at Spatially Distinct Locations

ICLR 2019 · Tobias Hinz, Stefan Heinrich, Stefan Wermter ·

Recent improvements to Generative Adversarial Networks (GANs) have made it possible to generate realistic images in high resolution based on natural language descriptions such as image captions. Furthermore, conditional GANs allow us to control the image generation process through labels or even natural language descriptions. However, fine-grained control of the image layout, i.e. where in the image specific objects should be located, is still difficult to achieve. This is especially true for images that should contain multiple distinct objects at different spatial locations. We introduce a new approach which allows us to control the location of arbitrarily many objects within an image by adding an object pathway to both the generator and the discriminator. Our approach does not need a detailed semantic layout but only bounding boxes and the respective labels of the desired objects are needed. The object pathway focuses solely on the individual objects and is iteratively applied at the locations specified by the bounding boxes. The global pathway focuses on the image background and the general image layout. We perform experiments on the Multi-MNIST, CLEVR, and the more complex MS-COCO data set. Our experiments show that through the use of the object pathway we can control object locations within images and can model complex scenes with multiple objects at various locations. We further show that the object pathway focuses on the individual objects and learns features relevant for these, while the global pathway focuses on global image characteristics and the image background.

PDF Abstract ICLR 2019 PDF ICLR 2019 Abstract

Code

Add Remove Mark official

tohinz/multiple-objects-gan official

113

Tasks

Add Remove

Conditional Image Generation

Image Generation

Object

Text-to-Image Generation

Datasets

MS COCO

CLEVR

SHAPES

Results from the Paper

Edit

Ranked #64 on Text-to-Image Generation on MS COCO

Get a GitHub badge

Task	Dataset	Model	Metric Name	Metric Value	Global Rank	Benchmark
Text-to-Image Generation	MS COCO	StackGAN + OP	FID	55.30	# 66	Compare
Text-to-Image Generation	MS COCO	StackGAN + OP	Inception score	12.12	# 22	Compare
Text-to-Image Generation	MS COCO	AttnGAN + OP	FID	33.35	# 62	Compare
			Inception score	24.76	# 16	Compare
			SOA-C	25.46	# 4	Compare

Methods

Add Remove

No methods listed for this paper. Add relevant methods here

Edit Social Preview

Generating Multiple Objects at Spatially Distinct Locations

Code Edit Add Remove Mark official

Tasks Edit Add Remove

Datasets Edit

Results from the Paper Edit

Methods Edit Add Remove

Code

Add Remove Mark official

Tasks

Add Remove

Datasets

Results from the Paper

Edit

Methods

Add Remove