资源论文SUMO: UNBIASED ESTIMATION OF LOG MARGINALP ROBABILITY FOR LATENT VARIABLE MODELS

SUMO: UNBIASED ESTIMATION OF LOG MARGINALP ROBABILITY FOR LATENT VARIABLE MODELS

2020-01-02 | |  59 |   33 |   0

Abstract

The standard variational lower bounds used to train latent variable models produce biased estimates of most quantities of interest. We introduce an unbiased estimator of the log marginal likelihood and its gradients for latent variable models based on randomized truncation of infinite series. If parameterized by an encoder-decoder architecture, the parameters of the encoder can be optimized to minimize its variance of this estimator. We show that models trained using our estimator give better test-set likelihoods than a standard importance-sampling based approach for the same average computational cost. This estimator also allows use of latent variable models for tasks where unbiased estimators, rather than marginal likelihood lower bounds, are preferred, such as minimizing reverse KL divergences and estimating score functions.

上一篇:CAUSAL DISCOVERY WITH REINFORCEMENT LEARNING

下一篇:SAMPLE EFFICIENT POLICY GRADIENT METHODSWITH RECURSIVE VARIANCE REDUCTION

用户评价
全部评价

热门资源

  • Learning to Predi...

    Much of model-based reinforcement learning invo...

  • Stratified Strate...

    In this paper we introduce Stratified Strategy ...

  • The Variational S...

    Unlike traditional images which do not offer in...

  • A Mathematical Mo...

    Direct democracy, where each voter casts one vo...

  • Rating-Boosted La...

    The performance of a recommendation system reli...