Skip to content
Open access

Enhancing Image Generation with GANs: The Role of Mutual Information in Optimizing Generative Models

Aug 2026 · Telematika · Vol 19, pp. 143-155 · 0 citations

TL;DR

It is demonstrated that mutual information regularization contributes to improved training stability, controllable generation, and more interpretable latent representations in GAN-based image generation.

Abstract

Generative Adversarial Networks (GANs) have become a prominent approach for image generation; however, they often suffer from training instability, mode collapse, and limited controllability of generated outputs. This study investigates the role of mutual information in improving generative modeling through a comparative analysis of Vanilla GAN, Conditional GAN (CGAN), and InfoGAN. Experiments were conducted using two image datasets with different levels of complexity, namely MNIST and Anime Face, under comparable training configurations. The evaluation focused on training behavior, convergence characteristics, generated image quality, and latent representation learning. The results revealed notable differences among the evaluated models. Vanilla GAN exhibited unstable convergence behavior at higher training epochs, while CGAN provided conditional control over generated outputs but did not fully mitigate training instability. In contrast, InfoGAN maintained more balanced generator and discriminator loss dynamics and produced visually consistent outputs across both datasets. Furthermore, latent code manipulation experiments showed that InfoGAN learned more structured and interpretable latent representations, enabling controllable feature variation in generated images. These findings indicate that incorporating mutual information improves representation learning, controllability, and training stability in GAN-based image generation. This study provides an empirical comparison of GAN, CGAN, and InfoGAN under a unified experimental framework and demonstrates that mutual information regularization contributes to improved training stability, controllable generation, and more interpretable latent representations. The findings highlight the potential of information-theoretic regularization for enhancing generative modeling performance.

Read PDF

Similar papers

#diffusion models Review Open access Sep 2026

Exploring the evolution of generative adversarial network architectures for text to image synthesis a comprehensive review

A structured analytical perspective on the evolution of GAN-based T2I synthesis is provided, identifies key design trade-offs, and outlines open challenges for future research.

Shreedatta S. Sawant, S. Kaliraj, S. Raghavendra et al. · 0 citations
Review Aug 2026

A Comprehensive Review of Generative Adversarial Networks in Deep Learning: Emerging Applications and Hybrid Architectures

Various strategies and improvements to enhance GANs stability and performance are examined, including hybrid architectures that integrate GANs with other deep learning models and practical utility in domain‐specific expert systems.

Kashif Iqbal, Xue Yu, Atifa Rafique et al. · 0 citations
#diffusion models Open access Sep 2026

A Systematic Comparison and Fusion of Generative Adversarial Networks and Diffusion Models for Image Generation

Generative Adversarial Networks (GANs) and diffusion models are two mainstream approaches in image generation. Although their underlying principles differ, both can be unified under the framework of particle models (PMs). GANs enable fast generation but suffer from unstable training and mode collapse, whereas diffusion...

Zhe-Ping Ding · 0 citations
#artificial intelligence Preprint Aug 2026

GAN-Diff : Coupling Pretrained WGAN-GP Features with Conditional Diffusion U-Nets

A hybrid GAN-guided diffusion framework that uses a pretrained Wasserstein GAN with gradient penalty (WGAN-GP) as a feature prior for conditional diffusion-based image restoration that consistently improves the quality of both degraded and low-resolution images.

Saif Ahmed, A. Galib, S. Antu et al. · 0 citations
Open access Aug 2026

Bridging Adversarial and Collaborative Learning for AI-Generated Image Quality Assessment

AI-generated image quality assessment (AIGIQA) requires jointly reasoning about perceptual fidelity and prompt alignment, two quality dimensions that are often treated as independent in existing AIGIQA models. However, by re-examining human ratings, we uncover a previously overlooked phenomenon: the two dimensions are...

Bao-Liang Chen, Qing Lin, Si-Jie Mai · 0 citations

with Pretrained Generative Model for Self-Supervised Learning

GenView is presented, a controllable framework that augments the diversity of positive views leveraging the power of pretrained generative models while preserving semantics, and an adaptive view generation method that dynamically adjusts the noise level in sampling to ensure the preservation of essential semantic meani...

Xiao-Jie Li, Yibo Yang, Xiang-Tai Li et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.