GFlowOut: Dropout with Generative Flow Networks

Liu, Dianbo; Jain, Moksh; Dossou, Bonaventure; Shen, Qianli; Lahlou, Salem; Goyal, Anirudh; Malkin, Nikolay; Emezue, Chris; Zhang, Dinghuai; Hassen, Nadhir; Ji, Xu; Kawaguchi, Kenji; Bengio, Yoshua

Computer Science > Machine Learning

arXiv:2210.12928 (cs)

[Submitted on 24 Oct 2022 (v1), last revised 24 Jun 2023 (this version, v3)]

Title:GFlowOut: Dropout with Generative Flow Networks

Authors:Dianbo Liu, Moksh Jain, Bonaventure Dossou, Qianli Shen, Salem Lahlou, Anirudh Goyal, Nikolay Malkin, Chris Emezue, Dinghuai Zhang, Nadhir Hassen, Xu Ji, Kenji Kawaguchi, Yoshua Bengio

View PDF

Abstract:Bayesian Inference offers principled tools to tackle many critical problems with modern neural networks such as poor calibration and generalization, and data inefficiency. However, scaling Bayesian inference to large architectures is challenging and requires restrictive approximations. Monte Carlo Dropout has been widely used as a relatively cheap way for approximate Inference and to estimate uncertainty with deep neural networks. Traditionally, the dropout mask is sampled independently from a fixed distribution. Recent works show that the dropout mask can be viewed as a latent variable, which can be inferred with variational inference. These methods face two important challenges: (a) the posterior distribution over masks can be highly multi-modal which can be difficult to approximate with standard variational inference and (b) it is not trivial to fully utilize sample-dependent information and correlation among dropout masks to improve posterior estimation. In this work, we propose GFlowOut to address these issues. GFlowOut leverages the recently proposed probabilistic framework of Generative Flow Networks (GFlowNets) to learn the posterior distribution over dropout masks. We empirically demonstrate that GFlowOut results in predictive distributions that generalize better to out-of-distribution data, and provide uncertainty estimates which lead to better performance in downstream tasks.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI)
Cite as:	arXiv:2210.12928 [cs.LG]
	(or arXiv:2210.12928v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2210.12928

Submission history

From: Dianbo Liu Dr [view email]
[v1] Mon, 24 Oct 2022 03:00:01 UTC (996 KB)
[v2] Mon, 7 Nov 2022 08:30:49 UTC (1,006 KB)
[v3] Sat, 24 Jun 2023 02:53:49 UTC (1,532 KB)

Computer Science > Machine Learning

Title:GFlowOut: Dropout with Generative Flow Networks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:GFlowOut: Dropout with Generative Flow Networks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators