Finding the global semantic representation in GAN through Frechet Mean
October 11, 2022 Β· Declared Dead Β· π International Conference on Learning Representations
"No code URL or promise found in abstract"
Evidence collected by the PWNC Scanner
Authors
Jaewoong Choi, Geonho Hwang, Hyunsoo Cho, Myungjoo Kang
arXiv ID
2210.05509
Category
cs.CV: Computer Vision
Citations
2
Venue
International Conference on Learning Representations
Last Checked
5 months ago
Abstract
The ideally disentangled latent space in GAN involves the global representation of latent space with semantic attribute coordinates. In other words, considering that this disentangled latent space is a vector space, there exists the global semantic basis where each basis component describes one attribute of generated images. In this paper, we propose an unsupervised method for finding this global semantic basis in the intermediate latent space in GANs. This semantic basis represents sample-independent meaningful perturbations that change the same semantic attribute of an image on the entire latent space. The proposed global basis, called FrΓ©chet basis, is derived by introducing FrΓ©chet mean to the local semantic perturbations in a latent space. FrΓ©chet basis is discovered in two stages. First, the global semantic subspace is discovered by the FrΓ©chet mean in the Grassmannian manifold of the local semantic subspaces. Second, FrΓ©chet basis is found by optimizing a basis of the semantic subspace via the FrΓ©chet mean in the Special Orthogonal Group. Experimental results demonstrate that FrΓ©chet basis provides better semantic factorization and robustness compared to the previous methods. Moreover, we suggest the basis refinement scheme for the previous methods. The quantitative experiments show that the refined basis achieves better semantic factorization while constrained on the same semantic subspace given by the previous method.
Community Contributions
Found the code? Know the venue? Think something is wrong? Let us know!
π Similar Papers
In the same crypt β Computer Vision
π
π
Old Age
π
π
Old Age
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
π
π
Old Age
SSD: Single Shot MultiBox Detector
π
π
Old Age
Squeeze-and-Excitation Networks
π
π
Old Age
Fast R-CNN
π
π
Old Age
Grad-CAM: Visual Explanations from Deep Networks via Gradient-based Localization
Died the same way β π» Ghosted
R.I.P.
π»
Ghosted
Federated Learning: Strategies for Improving Communication Efficiency
R.I.P.
π»
Ghosted
In-Datacenter Performance Analysis of a Tensor Processing Unit
R.I.P.
π»
Ghosted
Deep Convolutional Neural Networks for Computer-Aided Detection: CNN Architectures, Dataset Characteristics and Transfer Learning
R.I.P.
π»
Ghosted