PPO Agent playing Pyramids

This is a trained model of a PPO agent playing Pyramids using the Unity ML-Agents Library.

This model was trained as part of the Hugging Face Deep Reinforcement Learning Course, Unit 5.

Environment

ML-Agents-Pyramids

Algorithm

Proximal Policy Optimization (PPO)

Library

Unity ML-Agents

Model

The trained agent is provided as an ONNX model.

Downloads last month
6
Video Preview
loading

Evaluation results