Semantically Multi-Modal Image Synthesis

Zhu, Zhen; Xu, Zhiliang; You, Ansheng; Bai, Xiang

doi:10.1109/CVPR42600.2020.00551

Semantically Multi-Modal Image Synthesis

Zhen Zhu, Zhiliang Xu, Ansheng You, Xiang Bai

CVPR 2020

doi:10.1109/CVPR42600.2020.00551 /cvpr/2020/zhu2020cvpr-semantically/

Abstract

In this paper, we focus on semantically multi-modal image synthesis (SMIS) task, namely, generating multi-modal images at the semantic level. Previous work seeks to use multiple class-specific generators, constraining its usage in datasets with a small number of classes. We instead propose a novel Group Decreasing Network (GroupDNet) that leverages group convolutions in the generator and progressively decreases the group numbers of the convolutions in the decoder. Consequently, GroupDNet is armed with much more controllability on translating semantic labels to natural images and has plausible high-quality yields for datasets with many classes. Experiments on several challenging datasets demonstrate the superiority of GroupDNet on performing the SMIS task. We also show that GroupDNet is capable of performing a wide range of interesting synthesis applications. Codes and models are available at: https://github.com/Seanseattle/SMIS.

PDF CVPR Semantic Scholar

Cite

Text

Zhu et al. "Semantically Multi-Modal Image Synthesis." Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020. doi:10.1109/CVPR42600.2020.00551

Markdown

[Zhu et al. "Semantically Multi-Modal Image Synthesis." Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2020.](https://mlanthology.org/cvpr/2020/zhu2020cvpr-semantically/) doi:10.1109/CVPR42600.2020.00551

BibTeX

@inproceedings{zhu2020cvpr-semantically,
  title     = {{Semantically Multi-Modal Image Synthesis}},
  author    = {Zhu, Zhen and Xu, Zhiliang and You, Ansheng and Bai, Xiang},
  booktitle = {Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition},
  year      = {2020},
  doi       = {10.1109/CVPR42600.2020.00551},
  url       = {https://mlanthology.org/cvpr/2020/zhu2020cvpr-semantically/}
}