Testing the Lottery Ticket Hypothesis by iteratively pruning networks to find sparse, trainable subnetworks.
A policy gradient agent trained from raw pixels to learn how to play Pong through self-play.
An agent that plays optimal Tic-Tac-Toe by modeling the game as a Markov Decision Process and solving for the best policy.