Image Generation

1990 papers with code • 85 benchmarks • 67 datasets

Image Generation (synthesis) is the task of generating new images from an existing dataset.

Unconditional generation refers to generating samples unconditionally from the dataset, i.e. $p(y)$
Conditional image generation (subtask) refers to generating samples conditionally from the dataset, based on a label, i.e. $p(y|x)$.

In this section, you can find state-of-the-art leaderboards for unconditional generation. For conditional generation, and other types of image generations, refer to the subtasks.

( Image credit: StyleGAN )

Benchmarks

Add a Result

These leaderboards are used to track progress in Image Generation

Dataset	Best Model	Compare
CIFAR-10	StyleSAN-XL	See all
ImageNet 64x64	RIN	See all
ImageNet 256x256	ViT-XL/2 with limited Interval Guidance	See all
FFHQ 256 x 256	StyleSAN-XL	See all
CelebA 64x64	DDPM-IP	See all
LSUN Bedroom 256 x 256	Diffusion ProjectedGAN	See all
ImageNet 32x32	StyleGAN-XL	See all
STL-10	Diffusion ProjectedGAN	See all
LSUN Churches 256 x 256	Projected GAN	See all
ImageNet 512x512	EDM2-XXL	See all
FFHQ 1024 x 1024	StyleSAN-XL	See all
CelebA 256x256	Efficient-VDVAE	See all
ImageNet 128x128	VDM++	See all
CelebA-HQ 256x256	RDM	See all
FFHQ-U	Alias-Free-R	See all
MNIST	Locally Masked PixelCNN (8 orders)	See all
CelebA-HQ 1024x1024	RDM	See all
Binarized MNIST	CR-NVAE	See all
LSUN Cat 256 x 256	Vision-aided GAN	See all
CelebA-HQ 128x128	U-Net GAN	See all
CIFAR-100	LeCAM (StyleGAN2 + ADA)	See all
AFHQV2	Polarity-StyleGAN3	See all
AFHQ Cat	Vision-aided GAN	See all
LSUN Horse 256 x 256	Vision-aided GAN	See all
CLEVR	Projected GAN	See all
Cityscapes	Projected GAN	See all
AFHQ Dog	Projected GAN	See all
Fashion-MNIST	PAE	See all
CelebA 128x128	U-Net GAN	See all
AFHQ Wild	Vision-aided GAN	See all
Places50	SinDiffusion	See all
CUB 128 x 128	Projected GAN	See all
Stanford Dogs	Projected GAN	See all
Stanford Cars	Projected GANs	See all
Pokemon 256x256	StyleGAN-XL	See all
VizDoom	GAUDI	See all
Replica	GAUDI	See all
VLN-CE	GAUDI	See all
ARKitScenes	GAUDI	See all
CAT 256x256	StyleGAN2 + DA + RLC (Ours)	See all
ADE-Indoor	Projected GAN	See all
Stacked MNIST	VAEBM	See all
CelebA-HQ 64x64	VAEBM	See all
CIFAR-10 (20% data)	DiffAugment-CR-BigGAN	See all
CIFAR-10 (10% data)	DiffAugment-StyleGAN2	See all
LSUN Bedroom	StyleGAN	See all
FFHQ 512 x 512	StyleSAN-XL	See all
FFHQ 128 x 128	DDPM-IP	See all
ObjectsRoom	GENESIS-V2	See all
ShapeStacks	GENESIS-V2	See all
MetFaces-U	Alias-Free-R	See all
MetFaces	t-Stylegan3-ada (NVIDIA pre-trained)	See all
Pokemon 1024x1024	StyleGAN-XL	See all
Oxford 102 Flowers 256 x 256	Projected GAN	See all
LSUN Car 512 x 384	Polarity-StyleGAN2	See all
LSUN Bedroom 64 x 64	WGAN-GP + TT Update Rule	See all
LSUN Bedroom 128 x 128	LadaGAN	See all
RC-49	cDRE-F-cSP+RS	See all
iNaturalist 2019	StyeGAN2 + NoisyTwins	See all
Cityscapes-5K 256x512	SB-GAN	See all
Cityscapes-25K 256x512	SB-GAN	See all
Indian Celebs 256 x 256	MSG-StyleGAN	See all
LSUN Car 256 x 256	StyleGAN2	See all
Multi-dSprites	GENESIS	See all
GQN	GENESIS	See all
Landscapes 256 x 256	CIPS	See all
Satellite-Buildings 256 x 256	CIPS	See all
Satellite-Landscapes 256 x 256	CIPS	See all
Oxford 102 Flowers 128x128	QSNGAN	See all
25% ImageNet 128x128	LeCAM + DA	See all
LLVIP	pix2pix	See all
SDSS Galaxies	AstroDDPM	See all
NASA Perseverance	Stylegan2-ada	See all
1,078 People 3D Faces Collection Data	Sessiz çığlık	See all
LSUN	BigGAN + gSR	See all
CelebA-HQ 512x512	RDM	See all
CelebA	PR-BigGAN - Recall	See all
LSUN tower 64x64	DDPM-IP	See all
FFHQ 64x64 - 4x upscaling	PFGM++	See all
KMNIST	Spiking-Diffusion	See all
EMNIST-Letters	Spiking-Diffusion	See all
ImageNet 256x256 - 1 labeled data per class	DPT	See all
ImageNet 256x256 - 2 labeled data per class	DPT	See all
ImageNet 256x256 - 5 labeled data per class	DPT	See all
ImageNet 256x256 - 1% labeled data	DPT	See all

Show all 85 benchmarks

Collapse benchmarks

Libraries

Use these libraries to find Image Generation models and implementations

open-mmlab/mmgeneration

9 papers

1,803

faceonlive/ai-research

9 papers

161

eriklindernoren/PyTorch-GAN

6 papers

15,728

stability-ai/generative-models

5 papers

22,310

See all 8 libraries.

Datasets

Subtasks

Conditional Image Generation

3D-Aware Image Synthesis

Facial Inpainting

Layout-to-Image Generation

ROI-based image generation

Image Generation from Scene Graphs

Pose-Guided Image Generation

User Constrained Thumbnail Generation

Handwritten Word Generation

Chinese Landscape Painting Generation

person reposing

Infinite Image Generation

Multi class one-shot image synthesis

Single class few-shot image synthesis

Latest papers

Most implemented Social Latest No code

ANCHOR: LLM-driven News Subject Conditioning for Text-to-Image Synthesis

aashish2000/anchor • 15 Apr 2024

With Large Language Models (LLM) achieving success in language and commonsense reasoning tasks, we explore the ability of different LLMs to identify and understand key subjects from abstractive captions.

15 Apr 2024

Paper
Code

Semantic Approach to Quantifying the Consistency of Diffusion Model Image Generation

brinnaebent/semantic-consistency-score • • 12 Apr 2024

In this study, we identify the need for an interpretable, quantitative score of the repeatability, or consistency, of image generation in diffusion models.

12 Apr 2024

Paper
Code

CAT: Contrastive Adapter Training for Personalized Image Generation

faceonlive/ai-research • 11 Apr 2024

Finally, we mention the possibility of CAT in the aspects of multi-concept adapter and optimization.

161

11 Apr 2024

Paper
Code

Taming Stable Diffusion for Text to 360° Panorama Image Generation

faceonlive/ai-research • 11 Apr 2024

Generative models, e. g., Stable Diffusion, have enabled the creation of photorealistic images from text prompts.

161

11 Apr 2024

Paper
Code

Latent Guard: a Safety Framework for Text-to-image Generation

faceonlive/ai-research • 11 Apr 2024

Hence, we propose Latent Guard, a framework designed to improve safety measures in text-to-image generation.

161

11 Apr 2024

Paper
Code

Model-based Cleaning of the QUILT-1M Pathology Dataset for Text-Conditional Image Synthesis

deepmicroscopy/quiltcleaner • • 11 Apr 2024

The QUILT-1M dataset is the first openly available dataset containing images harvested from various online sources.

11 Apr 2024

Paper
Code

A Gauss-Newton Approach for Min-Max Optimization in Generative Adversarial Networks

faceonlive/ai-research • 10 Apr 2024

It modifies the Gauss-Newton method to approximate the min-max Hessian and uses the Sherman-Morrison inversion formula to calculate the inverse.

161

10 Apr 2024

Paper
Code

Deep Generative Data Assimilation in Multimodal Setting

faceonlive/ai-research • 10 Apr 2024

To our knowledge, our work is the first to apply deep generative framework for multimodal data assimilation using real-world datasets; an important step for building robust computational simulators, including the next-generation Earth system models.

161

10 Apr 2024

Paper
Code

StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion

faceonlive/ai-research • 9 Apr 2024

3) The story visualization and continuation models are trained and inferred independently, which is not user-friendly.

161

09 Apr 2024

Paper
Code

Hyperparameter-Free Medical Image Synthesis for Sharing Data and Improving Site-Specific Segmentation

faceonlive/ai-research • 9 Apr 2024

Sharing synthetic medical images is a promising alternative to sharing real images that can improve patient privacy and data security.

161

09 Apr 2024

Paper
Code

Image Generation

Benchmarks Add a Result

Libraries

Datasets

Subtasks

Latest papers

Content

Benchmarks

Add a Result