This model was trained for the Hugging Face Deep Reinforcement Learning course using a CleanRL-style PPO implementation in PyTorch.
-
Create images in seconds. No sign-up, no paywall, no setup.