Reinforcement Learning
ml-agents
TensorBoard
ONNX
Pyramids
deep-reinforcement-learning
ML-Agents-Pyramids
Instructions to use VLat3/ppo-Pyramids with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- ml-agents
How to use VLat3/ppo-Pyramids with ml-agents:
mlagents-load-from-hf --repo-id="VLat3/ppo-Pyramids" --local-dir="./download: string[]s"
- Notebooks
- Google Colab
- Kaggle
| { | |
| "Pyramids": { | |
| "checkpoints": [ | |
| { | |
| "steps": 499965, | |
| "file_path": "results/Pyramids Training/Pyramids/Pyramids-499965.onnx", | |
| "reward": -0.9992000460624695, | |
| "creation_time": 1783027643.137093, | |
| "auxillary_file_paths": [ | |
| "results/Pyramids Training/Pyramids/Pyramids-499965.pt" | |
| ] | |
| }, | |
| { | |
| "steps": 999897, | |
| "file_path": "results/Pyramids Training/Pyramids/Pyramids-999897.onnx", | |
| "reward": 1.3898499757051468, | |
| "creation_time": 1783028815.9499588, | |
| "auxillary_file_paths": [ | |
| "results/Pyramids Training/Pyramids/Pyramids-999897.pt" | |
| ] | |
| }, | |
| { | |
| "steps": 1000006, | |
| "file_path": "results/Pyramids Training/Pyramids/Pyramids-1000006.onnx", | |
| "reward": 1.4314221955007977, | |
| "creation_time": 1783028816.034819, | |
| "auxillary_file_paths": [ | |
| "results/Pyramids Training/Pyramids/Pyramids-1000006.pt" | |
| ] | |
| } | |
| ], | |
| "final_checkpoint": { | |
| "steps": 1000006, | |
| "file_path": "results/Pyramids Training/Pyramids.onnx", | |
| "reward": 1.4314221955007977, | |
| "creation_time": 1783028816.034819, | |
| "auxillary_file_paths": [ | |
| "results/Pyramids Training/Pyramids/Pyramids-1000006.pt" | |
| ] | |
| } | |
| }, | |
| "metadata": { | |
| "stats_format_version": "0.3.0", | |
| "mlagents_version": "1.2.0.dev0", | |
| "torch_version": "2.8.0+cu128" | |
| } | |
| } |