physer's picture
Upload folder using huggingface_hub
ed7216b verified
|
Raw
History Blame Contribute Delete
1.29 kB
# rl-checkpoints/
Residual-PPO (rsl_rl) checkpoints and full logs for the EgoEngine refine
chunks that replay/MPC could not fix (`pending_rl` escalation, see
`docs/REFINE.md` §7 in the code repo).
Layout:
- `<triplet>/<seq_id>/chunk_<a>_<b>/` — local run: `logs/model_final.pt`
(rsl_rl checkpoint), tensorboard event files, `train_log*.txt`,
`eval_residual.json` + `eval_residual.png`, `chunk_start_state.npz`.
- `a800/<triplet>/<seq_id>/chunk_<a>_<b>_<tag>/` — remote 4xA800 batch runs
with the RL.md §5 improvement bundle (dense-after-violation, contact-or,
warm-start-mpc, feasibility curriculum; `v2_m05f15` = mass x0.5 /
friction 1.5 sweep variant).
Produced by (repo https://github.com/physercoe/egoengine-repro, see
`docs/RL.md` + `docs/A800_RL.md`):
```
OMNI_KIT_ACCEPT_EULA=YES .venv-sim/bin/python scripts/train_residual.py \
"<triplet>" <seq_id> --chunk <a>:<b> --dense-after-violation --contact-or \
--warm-start-mpc --curriculum-start-scale 2.0
# batch: scripts/run_rl_a800.sh
```
Honest status: the stage is delivered but under-budget — the canonical
knife-slip chunk `[40,60)` was NOT fixed (tool feasible_frac 0.0 before and
after RL, see `eval_residual.json`); candid negative result, analysis in
`docs/RL.md`.
License: MIT (our checkpoints/logs).