English

Multimodal hierarchical Variational AutoEncoders with Factor Analysis latent space

Machine Learning 2024-10-23 v3 Artificial Intelligence

Abstract

Purpose: Handling heterogeneous and mixed data types has become increasingly critical with the exponential growth in real-world databases. While deep generative models attempt to merge diverse data views into a common latent space, they often sacrifice interpretability, flexibility, and modularity. This study proposes a novel method to address these limitations by combining Variational AutoEncoders (VAEs) with a Factor Analysis latent space (FA-VAE). Methods: The proposed FA-VAE method employs multiple VAEs to learn a private representation for each heterogeneous data view in a continuous latent space. Information is shared between views using a low-dimensional latent space, generated via a linear projection matrix. This modular design creates a hierarchical dependency between private and shared latent spaces, allowing for the flexible addition of new views and conditioning of pre-trained models. Results: The FA-VAE approach facilitates cross-generation of data from different domains and enables transfer learning between generative models. This allows for effective integration of information across diverse data views while preserving their distinct characteristics. Conclusions: By overcoming the limitations of existing methods, the FA-VAE provides a more interpretable, flexible, and modular solution for managing heterogeneous data types. It offers a pathway to more efficient and scalable data-handling strategies, enhancing the potential for cross-domain data synthesis and model transferability.

Keywords

Cite

@article{arxiv.2207.09185,
  title  = {Multimodal hierarchical Variational AutoEncoders with Factor Analysis latent space},
  author = {Alejandro Guerrero-López and Carlos Sevilla-Salcedo and Vanessa Gómez-Verdejo and Pablo M. Olmos},
  journal= {arXiv preprint arXiv:2207.09185},
  year   = {2024}
}

Comments

21 pages main work, 2 pages supplementary, 14 figures

R2 v1 2026-06-25T01:02:47.923Z