A project done as part of Reinforcement Learning Lecture at New York University leveraging PPO and MPC.