PPO Agent playing Pyramids in Unity ML-Agents

This is a trained model of a PPO agent playing ML-Agents-Pyramids using Unity ML-Agents.

  • Mean Reward: 1.95 +/- 0.15
  • Result (mean - std): 1.80
Downloads last month
14
Video Preview
loading

Evaluation results