File size: 1,546 Bytes
a73ed65 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 | ---
license: other
tags:
- robotics
- fastwam
- flashwam
- place-cube
---
# pcbnew_flashwam_ft — FusedKV / RoPE-fixed FlashWAM, finetuned on place_cube_new (500 traj)
FlashWAM (`M1_FusedKV_RopeFixed`) action-video model, finetuned on the new
500-episode **place-cube-in-bowl** dataset ("pick and place new").
- **Model:** FasterWAM decoupled, `kv_source_mode: fused_kv`, `fixed_rope: true`
(video DiT Wan2.2-TI2V-5B backbone + 1-layer action DiT).
- **Data:** `place_cube_new_lerobot_v21` — 500 episodes / 51,579 frames, 10 Hz,
2 cameras (base + wrist), 7-dim delta action, 8-dim proprio (real gripper widths).
Task string: *"place the cube in the bowl"*.
- **Init:** finetuned from LIBERO checkpoint
`M1_FusedKV_RopeFixed_step021700` (new fused_kv format).
- **Training:** 30 epochs / 48,360 steps, 4×H200, global batch 32, lr 1e-4 cosine,
bf16. Final `loss=0.0802`, `loss_action=0.0045`.
## Checkpoints (`checkpoints/weights/`)
Weights-only checkpoints saved every 5 epochs:
| file | step | epoch |
|------|------|-------|
| `step_008060.pt` | 8,060 | 5 |
| `step_016120.pt` | 16,120 | 10 |
| `step_024180.pt` | 24,180 | 15 |
| `step_032240.pt` | 32,240 | 20 |
| `step_040300.pt` | 40,300 | 25 |
| `step_048360.pt` | 48,360 | 30 (final) |
Each file is ~10.13 GB (bf16 full model: fused video + action experts).
## Other files
- `config.yaml` — full training/model config for this run.
- `dataset_stats.json` — per-run action/state normalization stats (min/max), needed
at inference for de/normalization.
|