Spaces:
Sleeping
Sleeping
Sync results/TRAINING_RUN_PROOF.md with public WandB Report URL
Browse files
results/TRAINING_RUN_PROOF.md
CHANGED
|
@@ -10,7 +10,7 @@
|
|
| 10 |
|---|---|
|
| 11 |
| **HF Job ID** | `69ed0e59d70108f37acded4e` |
|
| 12 |
| **HF Job URL** | https://huggingface.co/jobs/pushpam14/69ed0e59d70108f37acded4e |
|
| 13 |
-
| **WandB run** (public) | https://wandb.ai/pushpamsubscriptions-inn/openenv-contract-guardian/
|
| 14 |
| **Trained adapter** (public) | https://huggingface.co/pushpam14/api-contract-validator-grpo-7b |
|
| 15 |
|
| 16 |
## Training configuration
|
|
@@ -68,7 +68,7 @@ Saved model to https://huggingface.co/pushpam14/api-contract-validator-grpo-7b
|
|
| 68 |
[INFO] uploading training_state.json -> .../training_artifacts/training_state.json
|
| 69 |
[INFO] done.
|
| 70 |
wandb: π View run grpo-7b-l4-300steps-v3 at:
|
| 71 |
-
https://wandb.ai/pushpamsubscriptions-inn/openenv-contract-guardian/
|
| 72 |
```
|
| 73 |
|
| 74 |
The full unfiltered log (3,534 lines, includes every per-step metric, every dependency download, every weight upload) is in [`training_full_log.txt`](training_full_log.txt) in this directory.
|
|
@@ -93,7 +93,7 @@ curl -sI https://huggingface.co/pushpam14/api-contract-validator-grpo-7b/resolve
|
|
| 93 |
# content-length: 138792
|
| 94 |
|
| 95 |
# WandB run is public β opens in any browser
|
| 96 |
-
open https://wandb.ai/pushpamsubscriptions-inn/openenv-contract-guardian/
|
| 97 |
```
|
| 98 |
|
| 99 |
WandB shows the full live training metrics β every step's reward, loss, gradient norm, KL divergence, and completion lengths. Cannot be faked or post-edited.
|
|
|
|
| 10 |
|---|---|
|
| 11 |
| **HF Job ID** | `69ed0e59d70108f37acded4e` |
|
| 12 |
| **HF Job URL** | https://huggingface.co/jobs/pushpam14/69ed0e59d70108f37acded4e |
|
| 13 |
+
| **WandB run** (public) | https://wandb.ai/pushpamsubscriptions-inn/openenv-contract-guardian/reports/Enterprise-Contract-Guardian-GRPO-training-Qwen-7B-LoRA-300-steps---VmlldzoxNjY3MTAxMA?accessToken=3dhumexjta1umyk04rq6dx47iww4t25utt3j0x7063b7pvzzibp8jah29grhlwpb |
|
| 14 |
| **Trained adapter** (public) | https://huggingface.co/pushpam14/api-contract-validator-grpo-7b |
|
| 15 |
|
| 16 |
## Training configuration
|
|
|
|
| 68 |
[INFO] uploading training_state.json -> .../training_artifacts/training_state.json
|
| 69 |
[INFO] done.
|
| 70 |
wandb: π View run grpo-7b-l4-300steps-v3 at:
|
| 71 |
+
https://wandb.ai/pushpamsubscriptions-inn/openenv-contract-guardian/reports/Enterprise-Contract-Guardian-GRPO-training-Qwen-7B-LoRA-300-steps---VmlldzoxNjY3MTAxMA?accessToken=3dhumexjta1umyk04rq6dx47iww4t25utt3j0x7063b7pvzzibp8jah29grhlwpb
|
| 72 |
```
|
| 73 |
|
| 74 |
The full unfiltered log (3,534 lines, includes every per-step metric, every dependency download, every weight upload) is in [`training_full_log.txt`](training_full_log.txt) in this directory.
|
|
|
|
| 93 |
# content-length: 138792
|
| 94 |
|
| 95 |
# WandB run is public β opens in any browser
|
| 96 |
+
open https://wandb.ai/pushpamsubscriptions-inn/openenv-contract-guardian/reports/Enterprise-Contract-Guardian-GRPO-training-Qwen-7B-LoRA-300-steps---VmlldzoxNjY3MTAxMA?accessToken=3dhumexjta1umyk04rq6dx47iww4t25utt3j0x7063b7pvzzibp8jah29grhlwpb
|
| 97 |
```
|
| 98 |
|
| 99 |
WandB shows the full live training metrics β every step's reward, loss, gradient norm, KL divergence, and completion lengths. Cannot be faked or post-edited.
|