Only awesome is awesome. This is a curation, not a collection.
This list is about what GANs are used for. General GAN papers whose contribution is the model
itself (DCGAN, BEGAN, WGAN, โฆ) are not included as entries โ only genuine landmarks make the
short list at the top. Pure theory / loss-function papers are out of scope; diffusion-only work is
out of scope (GAN+diffusion hybrids are fine).
๐ท Legend: ๐ paper ยท ๐ป code ยท ๐ project/demo ยท ๐ฌ video ยท ๐ blog
๐ค Contributions welcome โ see CONTRIBUTING.md.
๐ The landmark papers that I respect
| Paper |
Venue |
Year |
Links |
| Generative Adversarial Networks |
NeurIPS |
2014 |
๐ ๐ป |
| Unsupervised Representation Learning with Deep Convolutional GANs (DCGAN) |
ICLR |
2016 |
๐ ๐ป |
| Improved Techniques for Training GANs |
NeurIPS |
2016 |
๐ ๐ป |
| BEGAN: Boundary Equilibrium Generative Adversarial Networks |
โ |
2017 |
๐ ๐ป |
| Training Generative Adversarial Networks with Limited Data (StyleGAN2-ADA) |
NeurIPS |
2020 |
๐ ๐ป |
| The GAN is dead; long live the GAN! A Modern GAN Baseline (R3GAN) |
NeurIPS |
2024 |
๐ ๐ป |
Contents
Click to expand the table of contents โ or just press command/ctrl + F to search for a keyword.
Applications using GANs
๐ค Font generation
| Title |
Venue |
Year |
Links |
| Learning Chinese Character Style with Conditional GAN (zi2zi) |
โ |
2017 |
๐ป ๐ |
| Artistic Glyph Image Synthesis via One-Stage Few-Shot Learning (AGIS-Net) |
SIGGRAPH Asia |
2019 |
๐ ๐ป |
| Attribute2Font: Creating Fonts You Want From Attributes |
SIGGRAPH |
2020 |
๐ ๐ป |
โ back to Contents
๐ด Anime character generation
| Title |
Venue |
Year |
Links |
| Towards the Automatic Anime Characters Creation with GANs |
โ |
2017 |
๐ |
| AnimeGANv2: Photo to Anime Style Transfer |
โ |
2021 |
๐ป |
โ back to Contents
๐ญ Face Stylization
| Title |
Venue |
Year |
Links |
| JoJoGAN: One Shot Face Stylization |
ECCV |
2022 |
๐ ๐ป |
โ back to Contents
๐ฎ Interactive Image generation
| Title |
Venue |
Year |
Links |
| Generative Visual Manipulation on the Natural Image Manifold (iGAN) |
ECCV |
2016 |
๐ ๐ป |
| Neural Photo Editing with Introspective Adversarial Networks |
ICLR |
2017 |
๐ ๐ป |
| Drag Your GAN: Interactive Point-based Manipulation on the Generative Image Manifold (DragGAN) |
SIGGRAPH |
2023 |
๐ ๐ป |
โ back to Contents
๐ง GAN Inversion and Latent Space Editing
| Title |
Venue |
Year |
Links |
| StyleCLIP: Text-Driven Manipulation of StyleGAN Imagery |
ICCV |
2021 |
๐ ๐ป |
| Encoding in Style: a StyleGAN Encoder for Image-to-Image Translation (pSp) |
CVPR |
2021 |
๐ ๐ป |
โ back to Contents
โ Text2Image (text to image)
| Title |
Venue |
Year |
Links |
| TAC-GAN: Text Conditioned Auxiliary Classifier GAN |
โ |
2017 |
๐ ๐ป |
| StackGAN: Text to Photo-realistic Image Synthesis with Stacked GANs |
ICCV |
2017 |
๐ ๐ป |
| Generative Adversarial Text to Image Synthesis |
ICML |
2016 |
๐ ๐ป ๐ป |
| Learning What and Where to Draw |
NeurIPS |
2016 |
๐ ๐ป |
| AttnGAN: Fine-Grained Text to Image Generation with Attentional GANs |
CVPR |
2018 |
๐ ๐ป |
| GigaGAN: Scaling up GANs for Text-to-Image Synthesis |
CVPR |
2023 |
๐ |
| StyleGAN-T: Unlocking the Power of GANs for Fast Large-Scale Text-to-Image Synthesis |
ICML |
2023 |
๐ ๐ป |
โ back to Contents
โก Adversarial Diffusion Distillation (GAN-hybrid)
GANs strike back in 2024โ2026: adversarial objectives distill slow diffusion samplers into one/few-step image and video generators โ where GANs are most alive in 2025โ2026.
| Title |
Venue |
Year |
Links |
| UFOGen: You Forward Once Large Scale Text-to-Image Generation via Diffusion GANs |
CVPR |
2024 |
๐ |
| Improved Distribution Matching Distillation for Fast Image Synthesis (DMD2) |
NeurIPS |
2024 |
๐ ๐ป |
| SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation |
arXiv |
2025 |
๐ ๐ป |
| Diffusion Adversarial Post-Training for One-Step Video Generation (Seaweed-APT) |
ICML |
2025 |
๐ ๐ |
| Autoregressive Adversarial Post-Training for Real-Time Interactive Video Generation (Seaweed-APT2) |
NeurIPS |
2025 |
๐ ๐ |
โ back to Contents
๐ง 3D Object generation
| Title |
Venue |
Year |
Links |
| Parametric 3D Exploration with Stacked Adversarial Networks (pix2vox) |
โ |
2016 |
๐ป ๐ฌ |
| Learning a Probabilistic Latent Space of Object Shapes via 3D Generative-Adversarial Modeling (3D-GAN) |
NeurIPS |
2016 |
๐ ๐ป ๐ฌ |
| 3D Shape Induction from 2D Views of Multiple Objects |
โ |
2016 |
๐ |
| Fully Convolutional Refined Auto-Encoding GANs for 3D Multi Object Scenes |
โ |
2017 |
๐ป ๐ |
โ back to Contents
๐ 3D-aware Image Synthesis
| Title |
Venue |
Year |
Links |
| Efficient Geometry-aware 3D Generative Adversarial Networks (EG3D) |
CVPR |
2022 |
๐ ๐ป |
โ back to Contents
โ Image Editing
| Title |
Venue |
Year |
Links |
| Invertible Conditional GANs for Image Editing (IcGAN) |
โ |
2016 |
๐ ๐ป |
| Image De-raining Using a Conditional GAN (ID-CGAN) |
โ |
2017 |
๐ ๐ป |
| DeblurGAN: Blind Motion Deblurring Using Conditional Adversarial Networks |
CVPR |
2018 |
๐ ๐ป |
โ back to Contents
๐ด Face Aging
| Title |
Venue |
Year |
Links |
| Age Progression/Regression by Conditional Adversarial Autoencoder (CAAE) |
CVPR |
2017 |
๐ ๐ป |
| CAN: Creative Adversarial Networks Generating "Art" |
โ |
2017 |
๐ |
| Face Aging with Conditional Generative Adversarial Networks |
โ |
2017 |
๐ |
โ back to Contents
๐บ Human Pose Estimation
| Title |
Venue |
Year |
Links |
| Joint Discriminative and Generative Learning for Person Re-identification (DG-Net) |
CVPR |
2019 |
๐ ๐ป ๐ฌ |
| Pose Guided Person Image Generation |
NeurIPS |
2017 |
๐ |
โ back to Contents
๐ฃ Talking Head and Face Reenactment
| Title |
Venue |
Year |
Links |
| First Order Motion Model for Image Animation |
NeurIPS |
2019 |
๐ ๐ป |
| A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild (Wav2Lip) |
ACM MM |
2020 |
๐ ๐ป |
โ back to Contents
๐ Virtual Try-On
| Title |
Venue |
Year |
Links |
| VITON-HD: High-Resolution Virtual Try-On via Misalignment-Aware Normalization |
CVPR |
2021 |
๐ ๐ป |
โ back to Contents
๐ Domain-transfer (e.g. style-transfer, pix2pix, sketch2image)
| Title |
Venue |
Year |
Links |
| Image-to-Image Translation with Conditional Adversarial Networks (pix2pix) |
CVPR |
2017 |
๐ ๐ป ๐ฌ |
| Unpaired Image-to-Image Translation using Cycle-Consistent Adversarial Networks (CycleGAN) |
ICCV |
2017 |
๐ ๐ป ๐ฌ |
| Learning to Discover Cross-Domain Relations with GANs (DiscoGAN) |
ICML |
2017 |
๐ ๐ป |
| StarGAN: Unified Multi-Domain Image-to-Image Translation |
CVPR |
2018 |
๐ ๐ป |
| StarGAN v2: Diverse Image Synthesis for Multiple Domains |
CVPR |
2020 |
๐ ๐ป |
| Multimodal Unsupervised Image-to-Image Translation (MUNIT) |
ECCV |
2018 |
๐ ๐ป |
| Unsupervised Creation of Parameterized Avatars |
โ |
2017 |
๐ |
| Unsupervised Cross-Domain Image Generation (DTN) |
ICLR |
2017 |
๐ |
| Precomputed Real-Time Texture Synthesis with Markovian GANs (MGANs) |
ECCV |
2016 |
๐ ๐ป |
| Pixel-Level Domain Transfer (PixelDTGAN) |
ECCV |
2016 |
๐ ๐ป |
| TextureGAN: Controlling Deep Image Synthesis with Texture Patches |
CVPR |
2018 |
๐ ๐ |
| Vincent AI Sketch Demo (NVIDIA, GTC Europe) |
โ |
2017 |
๐ ๐ฌ |
| Deep Photo Style Transfer |
CVPR |
2017 |
๐ ๐ป |
โ back to Contents
๐ฉน Image Inpainting (hole filling)
| Title |
Venue |
Year |
Links |
| Context Encoders: Feature Learning by Inpainting |
CVPR |
2016 |
๐ ๐ป |
| Semantic Image Inpainting with Perceptual and Contextual Losses |
CVPR |
2017 |
๐ ๐ป |
| Semi-Supervised Learning with Context-Conditional GANs |
โ |
2016 |
๐ |
| Free-Form Image Inpainting with Gated Convolution (DeepFill v2) |
ICCV |
2019 |
๐ ๐ป |
| Resolution-robust Large Mask Inpainting with Fourier Convolutions (LaMa) |
WACV |
2022 |
๐ ๐ป |
โ back to Contents
๐ช Image Blending
| Title |
Venue |
Year |
Links |
| GP-GAN: Towards Realistic High-Resolution Image Blending |
ACM MM |
2019 |
๐ ๐ป |
โ back to Contents
๐ Super-resolution
| Title |
Venue |
Year |
Links |
| Image Super-Resolution Through Deep Learning (srez) |
โ |
2016 |
๐ป |
| Photo-Realistic Single Image Super-Resolution Using a GAN (SRGAN) |
CVPR |
2017 |
๐ ๐ป |
| High-Quality Face Image Super-Resolution Using Conditional GANs |
โ |
2017 |
๐ |
| Analyzing Perception-Distortion Tradeoff (EPSR) |
ECCVW |
2018 |
๐ ๐ป |
| ESRGAN: Enhanced Super-Resolution Generative Adversarial Networks |
ECCVW |
2018 |
๐ ๐ป |
| Real-ESRGAN: Training Real-World Blind Super-Resolution with Pure Synthetic Data |
ICCVW |
2021 |
๐ ๐ป |
โ back to Contents
๐ผ High-resolution image generation (large-scale image)
| Title |
Venue |
Year |
Links |
| Generating Large Images from Latent Vectors |
โ |
2016 |
๐ป ๐ |
| Progressive Growing of GANs for Improved Quality, Stability, and Variation (PGGAN) |
ICLR |
2018 |
๐ ๐ป |
| Large Scale GAN Training for High Fidelity Natural Image Synthesis (BigGAN) |
ICLR |
2019 |
๐ |
| SinGAN: Learning a Generative Model from a Single Natural Image |
ICCV |
2019 |
๐ ๐ป |
| Analyzing and Improving the Image Quality of StyleGAN (StyleGAN2) |
CVPR |
2020 |
๐ ๐ป |
| Alias-Free Generative Adversarial Networks (StyleGAN3) |
NeurIPS |
2021 |
๐ ๐ป |
โ back to Contents
๐ก Adversarial Examples (Defense vs Attack)
| Title |
Venue |
Year |
Links |
| SafetyNet: Detecting and Rejecting Adversarial Examples Robustly |
ICCV |
2017 |
๐ |
| Adversarial Examples for Generative Models |
โ |
2017 |
๐ |
โ back to Contents
๐ Visual Saliency Prediction (attention prediction)
| Title |
Venue |
Year |
Links |
| SalGAN: Visual Saliency Prediction with Generative Adversarial Networks |
โ |
2017 |
๐ ๐ป |
โ back to Contents
๐ฏ Object Detection/Recognition
| Title |
Venue |
Year |
Links |
| Perceptual Generative Adversarial Networks for Small Object Detection |
CVPR |
2017 |
๐ |
| Adversarial Generation of Training Examples for Vehicle License Plate Recognition |
โ |
2017 |
๐ |
โ back to Contents
๐ฌ Video (generation/prediction)
| Title |
Venue |
Year |
Links |
| Deep Multi-Scale Video Prediction Beyond Mean Square Error |
ICLR |
2016 |
๐ ๐ป |
| Learning Temporal Coherence via Self-Supervision for GAN-based Video Generation (TecoGAN) |
SIGGRAPH |
2020 |
๐ ๐ป |
โ back to Contents
๐ Audio and Speech Synthesis
| Title |
Venue |
Year |
Links |
| Adversarial Audio Synthesis (WaveGAN) |
ICLR |
2019 |
๐ ๐ป |
| HiFi-GAN: GANs for Efficient and High Fidelity Speech Synthesis |
NeurIPS |
2020 |
๐ ๐ป |
| BigVGAN: A Universal Neural Vocoder with Large-Scale Training (v2 2024) |
ICLR |
2023 |
๐ ๐ป |
| Vocos: Closing the Gap between Time-domain and Fourier-based Neural Vocoders |
ICLR |
2024 |
๐ ๐ป |
| BemaGANv2: Discriminator Combination Strategies for GAN-based Vocoders in Long-Term Audio |
arXiv |
2025 |
๐ ๐ป |
โ back to Contents
๐ก Text and Sequence Generation
| Title |
Venue |
Year |
Links |
| SeqGAN: Sequence Generative Adversarial Nets with Policy Gradient |
AAAI |
2017 |
๐ ๐ป |
โ back to Contents
๐จ Anomaly Detection
| Title |
Venue |
Year |
Links |
| Unsupervised Anomaly Detection with GANs to Guide Marker Discovery (AnoGAN) |
IPMI |
2017 |
๐ |
โ back to Contents
๐งช Synthetic Data Generation
| Title |
Venue |
Year |
Links |
| Learning from Simulated and Unsupervised Images through Adversarial Training (SimGAN) |
CVPR |
2017 |
๐ ๐ป |
โ back to Contents
๐งฉ Others
| Title |
Venue |
Year |
Links |
| (Physics) Location-Aware Generative Adversarial Networks for Physics Synthesis |
โ |
2017 |
๐ ๐ป |
| (General) Spectral Normalization for Generative Adversarial Networks |
ICLR |
2018 |
๐ ๐ป |
โ back to Contents
Did not use GAN, but still interesting applications.
๐ง Real-time face reconstruction
| Title |
Venue |
Year |
Links |
| Model-based Deep Convolutional Face Autoencoder for Unsupervised Monocular Reconstruction (MoFA) |
ICCV |
2017 |
๐ ๐ป ๐ฌ |
โ back to Contents
๐ Super-resolution
| Title |
Venue |
Year |
Links |
| Learning to Simplify: Fully Convolutional Networks for Rough Sketch Cleanup |
SIGGRAPH |
2016 |
๐ ๐ฌ |
โ back to Contents
๐ผ Photorealistic Image generation (e.g. pix2pix, sketch2image)
| Title |
Venue |
Year |
Links |
| The Sketchy Database: Learning to Retrieve Badly Drawn Bunnies |
SIGGRAPH |
2016 |
๐ฌ |
| PatchMatch: A Randomized Correspondence Algorithm for Structural Image Editing |
SIGGRAPH |
2009 |
๐ ๐ป ๐ฌ |
โ back to Contents
๐บ Human Pose Estimation
| Title |
Venue |
Year |
Links |
| Knowledge-Guided Deep Fractal Neural Networks for Human Pose Estimation |
โ |
2017 |
๐ ๐ป |
โ back to Contents
๐ง 3D Object generation
| Title |
Venue |
Year |
Links |
| 3D-R2N2: A Unified Approach for Single and Multi-view 3D Object Reconstruction |
ECCV |
2016 |
๐ ๐ป |
โ back to Contents
๐ GAN tutorials with easy and simple example code for starters
โ back to Contents
Genuinely awesome, widely-used GitHub repos built on GANs โ usable code and tools, not just papers. Star counts update automatically.
Foundational model code (official / canonical)
โ back to Contents
๐งฐ Implementations of various types of GANs collection
โ back to Contents
๐ฐ Trendy AI-application Articles
โ back to Contents
Author
Minchul Shin, @nashory
Any recommendations to add to the list are welcome โ see CONTRIBUTING.md and feel free to make pull requests! :)