kangdawei commited on
Commit
97a871f
·
verified ·
1 Parent(s): 3e03734

Training in progress, step 500

Browse files
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:f702f57859188679c4525beb909f9898a8f6cee927fa0e1da6cc7364eeaa4937
3
  size 3554214752
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:ce70e595072ae102f8e20214f016aef4112c11e9526cf8880934d98c633bd879
3
  size 3554214752
reward_data/all_rewards.csv CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:802417baa349cee586941c622de92fefd916a4847e149b22e90fc93554499308
3
- size 271135598
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:7202faddb86dae855d2ec17f16cc4094763573bc166c1b2a48e1bb45b50e68e7
3
+ size 296789745
reward_plots/advantage_plot_step_450.png ADDED
reward_plots/advantage_plot_step_460.png ADDED
reward_plots/advantage_plot_step_470.png ADDED
reward_plots/advantage_plot_step_480.png ADDED
reward_plots/advantage_plot_step_490.png ADDED
reward_plots/reward_comparison_step_450.png ADDED
reward_plots/reward_comparison_step_460.png ADDED
reward_plots/reward_comparison_step_470.png ADDED
reward_plots/reward_comparison_step_480.png ADDED
reward_plots/reward_comparison_step_490.png ADDED