kangdawei commited on
Commit
8cd5121
·
verified ·
1 Parent(s): f9607e1

Training in progress, step 500

Browse files
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:0d94c38d1037e5f95ae18ba29df7dcf4790ccf610274dc856f289adcb124f2b3
3
  size 3554214752
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:b796276b315f5df3460d57188f922055eea7766535e4cea13b39e7779cb6ebed
3
  size 3554214752
reward_data/all_rewards.csv CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:258ed499daa896f9252b0187eca18ef9dd254a54c29b374013d999d42c2847bd
3
- size 373789979
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:8690a226dca7fa283473b5d1ba2c1424118937ccaf2c521b85571eb4835121c2
3
+ size 376655767
reward_plots/advantage_plot_step_450.png ADDED
reward_plots/advantage_plot_step_460.png ADDED
reward_plots/advantage_plot_step_470.png ADDED
reward_plots/advantage_plot_step_480.png ADDED
reward_plots/advantage_plot_step_490.png ADDED
reward_plots/reward_comparison_step_450.png ADDED
reward_plots/reward_comparison_step_460.png ADDED
reward_plots/reward_comparison_step_470.png ADDED
reward_plots/reward_comparison_step_480.png ADDED
reward_plots/reward_comparison_step_490.png ADDED