PaperScope
LIVE · 2026-10-06 05:40 UTC

Empirical Variational Autoencoder

Kaede Shiohara

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2610.06545 v1
Category
Submitted
2026-10-05

Abstract

We present Empirical Variational Autoencoder, a general generative framework for continuous-valued (i.e., non-vector-quantized) sequences. EVA is based on the evidence lower bound of the Variational Autoencoder (VAE) but learns autoregressive latent priors empirically from training data, which can be implemented only by an additional single linear layer on top of VAEs. By replacing the conventional standard-Gaussian constraint with the self-predicted priors, EVA significantly alleviates the latent distribution gap between prior and posterior which is typically observed in conventional VAEs, and leads to high-fidelity ancestral sampling for sequential data generation. Extensive experiments on image and sound synthesis demonstrate that EVA achieves competitive generation quality with autoregressive diffusion baselines despite its much faster inference time.

Comment: Project page: https://mapooon.github.io/EVAPage

arXiv abs page · PDF · same-day batch