Salem Lahlou

GFlowNet Foundations

Edward J Hu

Mo Tiwari

2021-11-17

ArXiv (prépublication)

GFlowNet Foundations

Edward J Hu

Mo Tiwari

2021-11-17

ArXiv (prépublication)

GFlowNet Foundations

Edward J Hu

Mo Tiwari

Generative Flow Networks (GFlowNets) have been introduced as a method to sample a diverse set of candidates in an active learning context, w… (voir plus)ith a training objective that makes them approximately sample in proportion to a given reward function. In this paper, we show a number of additional theoretical properties of GFlowNets. They can be used to estimate joint probability distributions and the corresponding marginal distributions where some variables are unspecified and, of particular interest, can represent distributions over composite objects like sets and graphs. GFlowNets amortize the work typically done by computationally expensive MCMC methods in a single but trained generative pass. They could also be used to estimate partition functions and free energies, conditional probabilities of supersets (supergraphs) given a subset (subgraph), as well as marginal distributions over all supersets (supergraphs) of a given set (graph). We introduce variations enabling the estimation of entropy and mutual information, sampling from a Pareto frontier, connections to reward-maximizing policies, and extensions to stochastic environments, continuous actions and modular energy functions.

2021-11-17

ArXiv (prépublication)

GFlowNet Foundations

Edward J Hu

Mo Tiwari

2021-11-17

ArXiv (prépublication)

GFlowNet Foundations

Edward J Hu

Mo Tiwari

2021-11-17

ArXiv (prépublication)

arxiv.org

Mastering Rate based Curriculum Learning

Lucas Willems

Salem Lahlou

Yoshua Bengio

2020-08-14

ArXiv (prépublication)

arxiv.org

BabyAI: A Platform to Study the Sample Efficiency of Grounded Language Learning

Maxime Chevalier-Boisvert

Allowing humans to interactively train artificial agents to understand language instructions is desirable for both practical and scientific … (voir plus)reasons, but given the poor data efficiency of the current learning methods, this goal may require substantial research efforts. Here, we introduce the BabyAI research platform to support investigations towards including humans in the loop for grounded language learning. The BabyAI platform comprises an extensible suite of 19 levels of increasing difficulty. The levels gradually lead the agent towards acquiring a combinatorially rich synthetic language which is a proper subset of English. The platform also provides a heuristic expert agent for the purpose of simulating a human teacher. We report baseline results and estimate the amount of human involvement that would be required to train a neural network-based agent on some of the BabyAI levels. We put forward strong evidence that current deep learning methods are not yet sufficiently sample efficient when it comes to learning a language with compositional properties.

2019-01-01

ICLR.cc/2019/Conference (poster)

openreview.net

BabyAI: First Steps Towards Grounded Language Learning With a Human In the Loop

Maxime Chevalier-Boisvert

Allowing humans to interactively train artificial agents to understand language instructions is desirable for both practical and scientific … (voir plus)reasons, but given the poor data efficiency of the current learning methods, this goal may require substantial research efforts. Here, we introduce the BabyAI research platform to support investigations towards including humans in the loop for grounded language learning. The BabyAI platform comprises an extensible suite of 19 levels of increasing difficulty. The levels gradually lead the agent towards acquiring a combinatorially rich synthetic language which is a proper subset of English. The platform also provides a heuristic expert agent for the purpose of simulating a human teacher. We report baseline results and estimate the amount of human involvement that would be required to train a neural network-based agent on some of the BabyAI levels. We put forward strong evidence that current deep learning methods are not yet sufficiently sample efficient when it comes to learning a language with compositional properties.

2018-10-18

arXiv.org (prépublication)

dblp.uni-trier.de

Conférence sur les politiques de l'IA de Mila

À l’avant-garde d’une nouvelle ère

Éclaireurs autochtones en IA

Salem Lahlou

Publications

Conférence sur les politiques de l'IA de Mila

À l’avant-garde d’une nouvelle ère

Éclaireurs autochtones en IA

Mots-clés populaires:

Salem Lahlou

Publications