Back to Search
Start Over
Learning Layout and Style Reconfigurable GANs for Controllable Image Synthesis.
- Source :
- IEEE Transactions on Pattern Analysis & Machine Intelligence; Sep2022, Vol. 44 Issue 9, p5070-5087, 18p
- Publication Year :
- 2022
-
Abstract
- With the remarkable recent progress on learning deep generative models, it becomes increasingly interesting to develop models for controllable image synthesis from reconfigurable structured inputs. This paper focuses on a recently emerged task, layout-to-image, whose goal is to learn generative models for synthesizing photo-realistic images from a spatial layout (i.e., object bounding boxes configured in an image lattice) and its style codes (i.e., structural and appearance variations encoded by latent vectors). This paper first proposes an intuitive paradigm for the task, layout-to-mask-to-image, which learns to unfold object masks in a weakly-supervised way based on an input layout and object style codes. The layout-to-mask component deeply interacts with layers in the generator network to bridge the gap between an input layout and synthesized images. Then, this paper presents a method built on Generative Adversarial Networks (GANs) for the proposed layout-to-mask-to-image synthesis with layout and style control at both image and object levels. The controllability is realized by a proposed novel Instance-Sensitive and Layout-Aware Normalization (ISLA-Norm) scheme. A layout semi-supervised version of the proposed method is further developed without sacrificing performance. In experiments, the proposed method is tested in the COCO-Stuff dataset and the Visual Genome dataset with state-of-the-art performance obtained. [ABSTRACT FROM AUTHOR]
- Subjects :
- GENERATIVE adversarial networks
COGNITIVE styles
DEEP learning
DNA-binding proteins
Subjects
Details
- Language :
- English
- ISSN :
- 01628828
- Volume :
- 44
- Issue :
- 9
- Database :
- Complementary Index
- Journal :
- IEEE Transactions on Pattern Analysis & Machine Intelligence
- Publication Type :
- Academic Journal
- Accession number :
- 158406170
- Full Text :
- https://doi.org/10.1109/TPAMI.2021.3078577