Image Generation

1996 papers with code • 85 benchmarks • 67 datasets

Image Generation (synthesis) is the task of generating new images from an existing dataset.

Unconditional generation refers to generating samples unconditionally from the dataset, i.e. $p(y)$
Conditional image generation (subtask) refers to generating samples conditionally from the dataset, based on a label, i.e. $p(y|x)$.

In this section, you can find state-of-the-art leaderboards for unconditional generation. For conditional generation, and other types of image generations, refer to the subtasks.

( Image credit: StyleGAN )

Benchmarks

Add a Result

These leaderboards are used to track progress in Image Generation

Dataset	Best Model	Compare
CIFAR-10	StyleSAN-XL	See all
ImageNet 64x64	RIN	See all
ImageNet 256x256	ViT-XL/2 with limited Interval Guidance	See all
FFHQ 256 x 256	StyleSAN-XL	See all
CelebA 64x64	DDPM-IP	See all
LSUN Bedroom 256 x 256	Diffusion ProjectedGAN	See all
ImageNet 32x32	StyleGAN-XL	See all
STL-10	Diffusion ProjectedGAN	See all
LSUN Churches 256 x 256	Projected GAN	See all
ImageNet 512x512	EDM2-XXL	See all
FFHQ 1024 x 1024	StyleSAN-XL	See all
CelebA 256x256	Efficient-VDVAE	See all
ImageNet 128x128	VDM++	See all
CelebA-HQ 256x256	RDM	See all
FFHQ-U	Alias-Free-R	See all
MNIST	Locally Masked PixelCNN (8 orders)	See all
CelebA-HQ 1024x1024	RDM	See all
Binarized MNIST	CR-NVAE	See all
LSUN Cat 256 x 256	Vision-aided GAN	See all
CelebA-HQ 128x128	U-Net GAN	See all
CIFAR-100	LeCAM (StyleGAN2 + ADA)	See all
AFHQV2	Polarity-StyleGAN3	See all
AFHQ Cat	Vision-aided GAN	See all
LSUN Horse 256 x 256	Vision-aided GAN	See all
CLEVR	Projected GAN	See all
Cityscapes	Projected GAN	See all
AFHQ Dog	Projected GAN	See all
Fashion-MNIST	PAE	See all
CelebA 128x128	U-Net GAN	See all
AFHQ Wild	Vision-aided GAN	See all
Places50	SinDiffusion	See all
CUB 128 x 128	Projected GAN	See all
Stanford Dogs	Projected GAN	See all
Stanford Cars	Projected GANs	See all
Pokemon 256x256	StyleGAN-XL	See all
VizDoom	GAUDI	See all
Replica	GAUDI	See all
VLN-CE	GAUDI	See all
ARKitScenes	GAUDI	See all
CAT 256x256	StyleGAN2 + DA + RLC (Ours)	See all
ADE-Indoor	Projected GAN	See all
Stacked MNIST	VAEBM	See all
CelebA-HQ 64x64	VAEBM	See all
CIFAR-10 (20% data)	DiffAugment-CR-BigGAN	See all
CIFAR-10 (10% data)	DiffAugment-StyleGAN2	See all
LSUN Bedroom	StyleGAN	See all
FFHQ 512 x 512	StyleSAN-XL	See all
FFHQ 128 x 128	DDPM-IP	See all
ObjectsRoom	GENESIS-V2	See all
ShapeStacks	GENESIS-V2	See all
MetFaces-U	Alias-Free-R	See all
MetFaces	t-Stylegan3-ada (NVIDIA pre-trained)	See all
Pokemon 1024x1024	StyleGAN-XL	See all
Oxford 102 Flowers 256 x 256	Projected GAN	See all
LSUN Car 512 x 384	Polarity-StyleGAN2	See all
LSUN Bedroom 64 x 64	WGAN-GP + TT Update Rule	See all
LSUN Bedroom 128 x 128	LadaGAN	See all
RC-49	cDRE-F-cSP+RS	See all
iNaturalist 2019	StyeGAN2 + NoisyTwins	See all
Cityscapes-5K 256x512	SB-GAN	See all
Cityscapes-25K 256x512	SB-GAN	See all
Indian Celebs 256 x 256	MSG-StyleGAN	See all
LSUN Car 256 x 256	StyleGAN2	See all
Multi-dSprites	GENESIS	See all
GQN	GENESIS	See all
Landscapes 256 x 256	CIPS	See all
Satellite-Buildings 256 x 256	CIPS	See all
Satellite-Landscapes 256 x 256	CIPS	See all
Oxford 102 Flowers 128x128	QSNGAN	See all
25% ImageNet 128x128	LeCAM + DA	See all
LLVIP	pix2pix	See all
SDSS Galaxies	AstroDDPM	See all
NASA Perseverance	Stylegan2-ada	See all
1,078 People 3D Faces Collection Data	Sessiz çığlık	See all
LSUN	BigGAN + gSR	See all
CelebA-HQ 512x512	RDM	See all
CelebA	PR-BigGAN - Recall	See all
LSUN tower 64x64	DDPM-IP	See all
FFHQ 64x64 - 4x upscaling	PFGM++	See all
KMNIST	Spiking-Diffusion	See all
EMNIST-Letters	Spiking-Diffusion	See all
ImageNet 256x256 - 1 labeled data per class	DPT	See all
ImageNet 256x256 - 2 labeled data per class	DPT	See all
ImageNet 256x256 - 5 labeled data per class	DPT	See all
ImageNet 256x256 - 1% labeled data	DPT	See all

Show all 85 benchmarks

Collapse benchmarks

Libraries

Use these libraries to find Image Generation models and implementations

open-mmlab/mmgeneration

9 papers

1,805

faceonlive/ai-research

9 papers

181

eriklindernoren/PyTorch-GAN

6 papers

15,737

stability-ai/generative-models

5 papers

22,349

See all 8 libraries.

Datasets

Subtasks

Conditional Image Generation

3D-Aware Image Synthesis

Facial Inpainting

Layout-to-Image Generation

ROI-based image generation

Image Generation from Scene Graphs

Pose-Guided Image Generation

User Constrained Thumbnail Generation

Handwritten Word Generation

Chinese Landscape Painting Generation

person reposing

Infinite Image Generation

Multi class one-shot image synthesis

Single class few-shot image synthesis

Latest papers with no code

Most implemented Social Latest No code

Synthesizing Iris Images using Generative Adversarial Networks: Survey and Comparative Analysis

no code yet • 26 Apr 2024

In this paper, we present a comprehensive review of state-of-the-art GAN-based synthetic iris image generation techniques, evaluating their strengths and limitations in producing realistic and useful iris images that can be used for both training and testing iris recognition systems and presentation attack detectors.

Paper
Add Code

BlenderAlchemy: Editing 3D Graphics with Vision-Language Models

no code yet • 26 Apr 2024

Specifically, we design a vision-based edit generator and state evaluator to work together to find the correct sequence of actions to achieve the goal.

Paper
Add Code

MuseumMaker: Continual Style Customization without Catastrophic Forgetting

no code yet • 25 Apr 2024

To deal with catastrophic forgetting amongst past learned styles, we devise a dual regularization for shared-LoRA module to optimize the direction of model update, which could regularize the diffusion model from both weight and feature aspects, respectively.

Paper
Add Code

Conditional Distribution Modelling for Few-Shot Image Synthesis with Diffusion Models

no code yet • 25 Apr 2024

Few-shot image synthesis entails generating diverse and realistic images of novel categories using only a few example images.

Paper
Add Code

Sketch2Human: Deep Human Generation with Disentangled Geometry and Appearance Control

no code yet • 24 Apr 2024

This work presents Sketch2Human, the first system for controllable full-body human image generation guided by a semantic sketch (for geometry control) and a reference image (for appearance control).

Paper
Add Code

SkinGEN: an Explainable Dermatology Diagnosis-to-Generation Framework with Interactive Vision-Language Models

no code yet • 23 Apr 2024

With the continuous advancement of vision language models (VLMs) technology, remarkable research achievements have emerged in the dermatology field, the fourth most prevalent human disease category.

Paper
Add Code

From Parts to Whole: A Unified Reference Framework for Controllable Human Image Generation

no code yet • 23 Apr 2024

Addressing this, we introduce Parts2Whole, a novel framework designed for generating customized portraits from multiple reference images, including pose images and various aspects of human appearance.

Paper
Add Code

Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation

no code yet • 23 Apr 2024

Recent studies have demonstrated the exceptional potentials of leveraging human preference datasets to refine text-to-image generative models, enhancing the alignment between generated images and textual prompts.

Paper
Add Code

FINEMATCH: Aspect-based Fine-grained Image and Text Mismatch Detection and Correction

no code yet • 23 Apr 2024

To address this, we propose FineMatch, a new aspect-based fine-grained text and image matching benchmark, focusing on text and image mismatch detection and correction.

Paper
Add Code

ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning

no code yet • 23 Apr 2024

The rapid development of diffusion models has triggered diverse applications.

Paper
Add Code

Image Generation

Benchmarks Add a Result

Libraries

Datasets

Subtasks

Latest papers with no code

Content

Benchmarks

Add a Result