Dual conditional GAN based on external attention for semantic image synthesis

Gang LiuSchool of Computer Science, Hubei University of Technology, Wuhan, People’s Republic of China

https://orcid.org/0000-0002-8589-0457 View further author information

Qijun ZhouSchool of Computer Science, Hubei University of Technology, Wuhan, People’s Republic of ChinaCorrespondence[email protected]

https://orcid.org/0009-0006-0233-6378 View further author information

Xiaoxiao XieSchool of Computer Science, Hubei University of Technology, Wuhan, People’s Republic of China

https://orcid.org/0009-0003-6132-077X View further author information

Qingchen YuSchool of Computer Science, Hubei University of Technology, Wuhan, People’s Republic of China

https://orcid.org/0009-0009-4613-7874 View further author information

Abstract

Although the existing semantic image synthesis methods based on generative adversarial networks (GANs) have achieved great success, the quality of the generated images still cannot achieve satisfactory results. This is mainly caused by two reasons. One reason is that the information in the semantic layout is sparse. Another reason is that a single constraint cannot effectively control the position relationship between objects in the generated image. To address the above problems, we propose a dual-conditional GAN with based on an external attention for semantic image synthesis (DCSIS). In DCSIS, the adaptive normalization method uses the one-hot encoded semantic layout to generate the first latent space and the external attention uses the RGB encoded semantic layout to generate the second latent space. Two latent spaces control the shape of objects and the positional relationship between objects in the generated image. The graph attention (GAT) is added to the generator to strengthen the relationship between different categories in the generated image. A graph convolutional segmentation network (GSeg) is designed to learn information for each category. Experiments on several challenging datasets demonstrate the advantages of our method over existing approaches, regarding both visual quality and the representative evaluating criteria.

KEYWORDS:

Disclosure statement

No potential conflict of interest was reported by the author(s).

Additional information

Funding

The work described in this article is supported by the Hubei Province University Student Innovation and Entrepreneurship Training, Hubei University of Technology Graduate Research Innovation Project [grant number 4306.22019]. The work described in this paper was support by the National Natural Science Foundation of China [grant number 61300127].

Dual conditional GAN based on external attention for semantic image synthesis

Information for

Open access

Opportunities

Help and information

Dual conditional GAN based on external attention for semantic image synthesis

Abstract

Disclosure statement

Additional information

Funding

Related research

To cite this article:

Download citation

Information for

Open access

Opportunities

Help and information

Keep up to date

Your download is now in progress and you may close this window

Login or register to access this feature