
Sign up to save your podcasts
Or


Today we’re joined by Jason Gauci, a Software Engineering Manager at Facebook AI.
In our conversation with Jason, we explore their Reinforcement Learning platform, Re-Agent (Horizon). We discuss the role of decision making and game theory in the platform and the types of decisions they’re using Re-Agent to make, from ranking and recommendations to their eCommerce marketplace.
Jason also walks us through the differences between online/offline and on/off policy model training, and where Re-Agent sits in this spectrum. Finally, we discuss the concept of counterfactual causality, and how they ensure safety in the results of their models.
The complete show notes for this episode can be found at twimlai.com/go/448.
By Sam Charrington4.7
422422 ratings
Today we’re joined by Jason Gauci, a Software Engineering Manager at Facebook AI.
In our conversation with Jason, we explore their Reinforcement Learning platform, Re-Agent (Horizon). We discuss the role of decision making and game theory in the platform and the types of decisions they’re using Re-Agent to make, from ranking and recommendations to their eCommerce marketplace.
Jason also walks us through the differences between online/offline and on/off policy model training, and where Re-Agent sits in this spectrum. Finally, we discuss the concept of counterfactual causality, and how they ensure safety in the results of their models.
The complete show notes for this episode can be found at twimlai.com/go/448.

1,095 Listeners

173 Listeners

303 Listeners

347 Listeners

224 Listeners

205 Listeners

210 Listeners

305 Listeners

97 Listeners

525 Listeners

133 Listeners

93 Listeners

228 Listeners

632 Listeners

34 Listeners