REINFORCE Agent playing Pixelcopter-PLE-v0

This is a trained model of a REINFORCE agent playing Pixelcopter-PLE-v0 using PyTorch.

  • Mean Reward: 8.65 +/- 1.15
  • Result (mean - std): 7.50
Downloads last month
12
Video Preview
loading

Evaluation results