상세 보기
DuDGAN: Improving Class-Conditional GANs via Dual-Diffusion
- Yeom, Taesun;
- Gu, Chanhoe;
- Lee, Minhyeok
WEB OF SCIENCE
6SCOPUS
10초록
Class-conditional image generation using generative adversarial networks (GANs) has been investigated through various techniques; however, it continues to face challenges such as mode collapse, training instability, and low-quality output in cases of datasets with high intra-class variation. Furthermore, most GANs often converge in larger iterations, resulting in poor iteration efficacy in training procedures. While Diffusion-GAN has shown potential in generating realistic samples, it has a critical limitation in generating class-conditional samples. To overcome these limitations, we propose a novel approach for class-conditional image generation using GANs called DuDGAN, which incorporates a dual diffusion-based noise injection process. DuDGAN consists of three unique networks: a discriminator, a generator, and a classifier. During the training process, Gaussian-mixture noises are injected into the two noise-aware networks, the discriminator and the classifier, in distinct ways. This noisy data helps to prevent overfitting by gradually introducing more challenging tasks, leading to improved model performance. As a result, DuDGAN outperforms state-of-the-art conditional GAN models for image generation in terms of performance. We evaluated DuDGAN using the AFHQ, Food-101, and CIFAR-10 datasets and observed superior results across metrics such as FID, KID, Precision, and Recall score compared with comparison models, highlighting the effectiveness of proposed approach. Authors
키워드
- 제목
- DuDGAN: Improving Class-Conditional GANs via Dual-Diffusion
- 저자
- Yeom, Taesun; Gu, Chanhoe; Lee, Minhyeok
- 발행일
- 2024
- 유형
- Article
- 저널명
- IEEE Access
- 권
- 12
- 페이지
- 39651 ~ 39661
- 언어
- ENG
- 출판사
- Institute of Electrical and Electronics Engineers Inc.
- 발행국가
- 미국
- 분량
- 11 페이지
- ISSN
- P 2169-3536