Coupled Gradient Estimators for Discrete Latent Variables

Dong, Zhe
Mnih, Andriy
Tucker, George

Publication date

June 2021

Language

English

Abstract

Training models with discrete latent variables is challenging due to the high variance of unbiased gradient estimators. While low-variance reparameterization gradients of a continuous relaxation can provide an effective solution, a continuous relaxation is not always available or tractable. Dong et al. (2020) and Yin et al. (2020) introduced a performant estimator that does not rely on continuous relaxations; however, it is limited to binary random variables. We introduce a novel derivation of their estimator based on importance sampling and statistical couplings, which we extend to the categorical setting. Motivated by the construction of a stick-breaking coupling, we introduce gradient estimators based on reparameterizing categorical vari...

Extracted data

We use cookies to provide a better user experience.

Data Protection

Coupled Gradient Estimators for Discrete Latent Variables

Abstract

Extracted data

Coupled Gradient Estimators for Discrete Latent Variables

Abstract

Extracted data

Related items

Related items