This work imposes an entropy-based constraint on individual latent variables, showing that the entropy of a latent code upper-bounds the mutual information it carries about the generative factors of the data, and proposes a weight-filter method that exploits the slack of the soft constraint to prune low-entropy dimensions during downstream training.
Abstract
The usefulness of a variational autoencoder (VAE) depends on two properties of its latent space that are hard to obtain together: high encoding capacity in the individual latent variables, and a low-dimensional, disentangled organization of those variables. Weakening the Kullback-Leibler regularization raises capacity but degrades disentanglement, while strengthening it prunes latent variables away entirely. We formulate VAE training as a soft-constrained optimization problem that addresses both. First, we impose an entropy-based constraint (EC) on individual latent variables, showing that the entropy of a latent code upper-bounds the mutual information it carries about the generative factors of the data. Second, we propose a weight-filter method that exploits the slack of the soft constraint to prune low-entropy dimensions during downstream training. On dSprites, the EC raises the aggregate latent-variable activation score by 43-62% over a vanilla VAE, attains the highest FactorVAE score among the \b{eta} \b{eta}-VAE variants (0.891 vs 0.847), and lowers reconstruction error by up to 38%. On MNIST, the weight filter reduces the latent dimensionality supplied to a downstream classifier from ten to two while holding accuracy above 90%, converging in 37% fewer epochs than the same procedure without the EC. We also find that low-entropy discrete factors tend to merge into a single latent variable, whereas high-entropy continuous factors are distributed across several.
The study demonstrates that the VAE possesses significant advantages in state compression and uncertainty modeling, but suffers from issues such as blurry image generation and oversimplified posterior distribution assumptions.
Xiao-Zhou Gao· Theoretical and Natural Scie...· 0 citations
Generative models for top- \(N\) recommendation have garnered significant attention, with Variational Autoencoder (VAE) emerging as a promising approach for modeling user preferences. Yet, traditional VAE-based models encounter two major challenges: simplistic priors may cause posterior collapse, resulting in ineffecti...
Xiao-Bo Guo, Shaoshuai Li, You-Ru Li et al.· ACM Transactions on Informat...· 0 citations
The Superposed Latent Autoencoder (SLAE) is introduced, which preserves high-capacity latent representations while sharing storage through learned superposition, and suggests a new principle for representation compression: instead of making every latent smaller, keep representations wide and let them share memory.
Quan-Ling Zhao, Jia-Ying Yang, Tian-Qi Zhang et al.· 0 citations
The variational graph autoencoder (VGAE) regularizes its posterior toward the prior with the Kullback-Leibler divergence, a choice inherited from the variational autoencoder rather than argued for. We introduce the generalized graph variational autoencoder (GGVA), which replaces that term with any member of the R\'enyi...
Kleyton da Costa, Bernardo Modenesi, I. Menezes et al.· 0 citations
The results suggest that architectural routing mechanisms may have negligible impact on core semantic understanding, with representational divergence confined to extreme structural margins.
Rithin Nagaraj, Rupa Laalasa Oruganti, Prerna Subhashchandra Kunder et al.· 0 citations
We present Empirical Variational Autoencoder, a general generative framework for continuous-valued (i.e., non-vector-quantized) sequences. EVA is based on the evidence lower bound of the Variational Autoencoder (VAE) but learns autoregressive latent priors empirically from training data, which can be implemented only b...
Kaede Shiohara· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.