On-line Adaptative Curriculum Learning for GANs

Doan, Thang; Monteiro, Joao; Albuquerque, Isabela; Mazoure, Bogdan; Durand, Audrey; Pineau, Joelle; Hjelm, R Devon

Computer Science > Machine Learning

arXiv:1808.00020 (cs)

[Submitted on 31 Jul 2018 (v1), last revised 11 Mar 2019 (this version, v6)]

Title:On-line Adaptative Curriculum Learning for GANs

Authors:Thang Doan, Joao Monteiro, Isabela Albuquerque, Bogdan Mazoure, Audrey Durand, Joelle Pineau, R Devon Hjelm

View PDF

Abstract:Generative Adversarial Networks (GANs) can successfully approximate a probability distribution and produce realistic samples. However, open questions such as sufficient convergence conditions and mode collapse still persist. In this paper, we build on existing work in the area by proposing a novel framework for training the generator against an ensemble of discriminator networks, which can be seen as a one-student/multiple-teachers setting. We formalize this problem within the full-information adversarial bandit framework, where we evaluate the capability of an algorithm to select mixtures of discriminators for providing the generator with feedback during learning. To this end, we propose a reward function which reflects the progress made by the generator and dynamically update the mixture weights allocated to each discriminator. We also draw connections between our algorithm and stochastic optimization methods and then show that existing approaches using multiple discriminators in literature can be recovered from our framework. We argue that less expressive discriminators are smoother and have a general coarse grained view of the modes map, which enforces the generator to cover a wide portion of the data distribution support. On the other hand, highly expressive discriminators ensure samples quality. Finally, experimental results show that our approach improves samples quality and diversity over existing baselines by effectively learning a curriculum. These results also support the claim that weaker discriminators have higher entropy improving modes coverage. Keywords: multiple discriminators, curriculum learning, multiple resolutions discriminators, multi-armed bandits, generative adversarial networks, smooth discriminators, multi-discriminator gan training, multiple experts.

Comments:	Accepted to the Thirty-Third AAAI Conference On Artificial Intelligence, 2019 (Added 128x128 CelebA samples to the end of the appendix)
Subjects:	Machine Learning (cs.LG); Machine Learning (stat.ML)
Cite as:	arXiv:1808.00020 [cs.LG]
	(or arXiv:1808.00020v6 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.1808.00020
Journal reference:	Proceedings of 33rd AAAI Conference on Artificial Intelligence (AAAI 2019)

Submission history

From: Bogdan Mazoure [view email]
[v1] Tue, 31 Jul 2018 18:34:56 UTC (22,135 KB)
[v2] Fri, 7 Sep 2018 22:56:38 UTC (10,521 KB)
[v3] Wed, 12 Sep 2018 15:52:10 UTC (8,463 KB)
[v4] Wed, 14 Nov 2018 17:17:59 UTC (8,484 KB)
[v5] Wed, 12 Dec 2018 17:58:03 UTC (14,059 KB)
[v6] Mon, 11 Mar 2019 17:15:48 UTC (14,059 KB)

Computer Science > Machine Learning

Title:On-line Adaptative Curriculum Learning for GANs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:On-line Adaptative Curriculum Learning for GANs

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators