Robotics
LeRobot
Safetensors
lingbot_va
Grigorij commited on
Commit
cd9eb22
·
verified ·
1 Parent(s): 388060a

checkpoint 020000

Browse files
checkpoints/020000/pretrained_model/README.md ADDED
@@ -0,0 +1,204 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ library_name: peft
3
+ tags:
4
+ - base_model:adapter:lerobot/lingbot_va_libero_long
5
+ - lora
6
+ ---
7
+
8
+ # Model Card for Model ID
9
+
10
+ <!-- Provide a quick summary of what the model is/does. -->
11
+
12
+
13
+
14
+ ## Model Details
15
+
16
+ ### Model Description
17
+
18
+ <!-- Provide a longer summary of what this model is. -->
19
+
20
+
21
+
22
+ - **Developed by:** [More Information Needed]
23
+ - **Funded by [optional]:** [More Information Needed]
24
+ - **Shared by [optional]:** [More Information Needed]
25
+ - **Model type:** [More Information Needed]
26
+ - **Language(s) (NLP):** [More Information Needed]
27
+ - **License:** [More Information Needed]
28
+ - **Finetuned from model [optional]:** [More Information Needed]
29
+
30
+ ### Model Sources [optional]
31
+
32
+ <!-- Provide the basic links for the model. -->
33
+
34
+ - **Repository:** [More Information Needed]
35
+ - **Paper [optional]:** [More Information Needed]
36
+ - **Demo [optional]:** [More Information Needed]
37
+
38
+ ## Uses
39
+
40
+ <!-- Address questions around how the model is intended to be used, including the foreseeable users of the model and those affected by the model. -->
41
+
42
+ ### Direct Use
43
+
44
+ <!-- This section is for the model use without fine-tuning or plugging into a larger ecosystem/app. -->
45
+
46
+ [More Information Needed]
47
+
48
+ ### Downstream Use [optional]
49
+
50
+ <!-- This section is for the model use when fine-tuned for a task, or when plugged into a larger ecosystem/app -->
51
+
52
+ [More Information Needed]
53
+
54
+ ### Out-of-Scope Use
55
+
56
+ <!-- This section addresses misuse, malicious use, and uses that the model will not work well for. -->
57
+
58
+ [More Information Needed]
59
+
60
+ ## Bias, Risks, and Limitations
61
+
62
+ <!-- This section is meant to convey both technical and sociotechnical limitations. -->
63
+
64
+ [More Information Needed]
65
+
66
+ ### Recommendations
67
+
68
+ <!-- This section is meant to convey recommendations with respect to the bias, risk, and technical limitations. -->
69
+
70
+ Users (both direct and downstream) should be made aware of the risks, biases and limitations of the model. More information needed for further recommendations.
71
+
72
+ ## How to Get Started with the Model
73
+
74
+ Use the code below to get started with the model.
75
+
76
+ [More Information Needed]
77
+
78
+ ## Training Details
79
+
80
+ ### Training Data
81
+
82
+ <!-- This should link to a Dataset Card, perhaps with a short stub of information on what the training data is all about as well as documentation related to data pre-processing or additional filtering. -->
83
+
84
+ [More Information Needed]
85
+
86
+ ### Training Procedure
87
+
88
+ <!-- This relates heavily to the Technical Specifications. Content here should link to that section when it is relevant to the training procedure. -->
89
+
90
+ #### Preprocessing [optional]
91
+
92
+ [More Information Needed]
93
+
94
+
95
+ #### Training Hyperparameters
96
+
97
+ - **Training regime:** [More Information Needed] <!--fp32, fp16 mixed precision, bf16 mixed precision, bf16 non-mixed precision, fp16 non-mixed precision, fp8 mixed precision -->
98
+
99
+ #### Speeds, Sizes, Times [optional]
100
+
101
+ <!-- This section provides information about throughput, start/end time, checkpoint size if relevant, etc. -->
102
+
103
+ [More Information Needed]
104
+
105
+ ## Evaluation
106
+
107
+ <!-- This section describes the evaluation protocols and provides the results. -->
108
+
109
+ ### Testing Data, Factors & Metrics
110
+
111
+ #### Testing Data
112
+
113
+ <!-- This should link to a Dataset Card if possible. -->
114
+
115
+ [More Information Needed]
116
+
117
+ #### Factors
118
+
119
+ <!-- These are the things the evaluation is disaggregating by, e.g., subpopulations or domains. -->
120
+
121
+ [More Information Needed]
122
+
123
+ #### Metrics
124
+
125
+ <!-- These are the evaluation metrics being used, ideally with a description of why. -->
126
+
127
+ [More Information Needed]
128
+
129
+ ### Results
130
+
131
+ [More Information Needed]
132
+
133
+ #### Summary
134
+
135
+
136
+
137
+ ## Model Examination [optional]
138
+
139
+ <!-- Relevant interpretability work for the model goes here -->
140
+
141
+ [More Information Needed]
142
+
143
+ ## Environmental Impact
144
+
145
+ <!-- Total emissions (in grams of CO2eq) and additional considerations, such as electricity usage, go here. Edit the suggested text below accordingly -->
146
+
147
+ Carbon emissions can be estimated using the [Machine Learning Impact calculator](https://mlco2.github.io/impact#compute) presented in [Lacoste et al. (2019)](https://arxiv.org/abs/1910.09700).
148
+
149
+ - **Hardware Type:** [More Information Needed]
150
+ - **Hours used:** [More Information Needed]
151
+ - **Cloud Provider:** [More Information Needed]
152
+ - **Compute Region:** [More Information Needed]
153
+ - **Carbon Emitted:** [More Information Needed]
154
+
155
+ ## Technical Specifications [optional]
156
+
157
+ ### Model Architecture and Objective
158
+
159
+ [More Information Needed]
160
+
161
+ ### Compute Infrastructure
162
+
163
+ [More Information Needed]
164
+
165
+ #### Hardware
166
+
167
+ [More Information Needed]
168
+
169
+ #### Software
170
+
171
+ [More Information Needed]
172
+
173
+ ## Citation [optional]
174
+
175
+ <!-- If there is a paper or blog post introducing the model, the APA and Bibtex information for that should go in this section. -->
176
+
177
+ **BibTeX:**
178
+
179
+ [More Information Needed]
180
+
181
+ **APA:**
182
+
183
+ [More Information Needed]
184
+
185
+ ## Glossary [optional]
186
+
187
+ <!-- If relevant, include terms and calculations in this section that can help readers understand the model or model card. -->
188
+
189
+ [More Information Needed]
190
+
191
+ ## More Information [optional]
192
+
193
+ [More Information Needed]
194
+
195
+ ## Model Card Authors [optional]
196
+
197
+ [More Information Needed]
198
+
199
+ ## Model Card Contact
200
+
201
+ [More Information Needed]
202
+ ### Framework versions
203
+
204
+ - PEFT 0.20.0
checkpoints/020000/pretrained_model/adapter_config.json ADDED
@@ -0,0 +1,48 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "alora_invocation_tokens": null,
3
+ "alpha_pattern": {},
4
+ "arrow_config": null,
5
+ "auto_mapping": {
6
+ "base_model_class": "LingBotVAPolicy",
7
+ "parent_library": "lerobot.policies.lingbot_va.modeling_lingbot_va"
8
+ },
9
+ "base_model_name_or_path": "lerobot/lingbot_va_libero_long",
10
+ "bias": "none",
11
+ "corda_config": null,
12
+ "ensure_weight_tying": false,
13
+ "eva_config": null,
14
+ "exclude_modules": null,
15
+ "fan_in_fan_out": false,
16
+ "inference_mode": true,
17
+ "init_lora_weights": true,
18
+ "layer_replication": null,
19
+ "layers_pattern": null,
20
+ "layers_to_transform": null,
21
+ "loftq_config": {},
22
+ "lora_alpha": 16,
23
+ "lora_bias": false,
24
+ "lora_dropout": 0.0,
25
+ "lora_ga_config": null,
26
+ "megatron_config": null,
27
+ "megatron_core": "megatron.core",
28
+ "modules_to_save": null,
29
+ "monteclora_config": null,
30
+ "peft_type": "LORA",
31
+ "peft_version": "0.20.0",
32
+ "qalora_group_size": 16,
33
+ "r": 16,
34
+ "rank_pattern": {},
35
+ "revision": null,
36
+ "target_modules": [
37
+ "to_q",
38
+ "to_v"
39
+ ],
40
+ "target_parameters": null,
41
+ "task_type": null,
42
+ "trainable_token_indices": null,
43
+ "use_bdlora": null,
44
+ "use_dora": false,
45
+ "use_qalora": false,
46
+ "use_rslora": false,
47
+ "velora_config": null
48
+ }
checkpoints/020000/pretrained_model/adapter_model.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:3aded1c2e636c749b375184c7835d4e7e059e12056349d6887637bb045340187
3
+ size 47218168
checkpoints/020000/pretrained_model/config.json ADDED
@@ -0,0 +1,100 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "type": "lingbot_va",
3
+ "n_obs_steps": 1,
4
+ "input_features": {
5
+ "observation.images.image": {
6
+ "type": "VISUAL",
7
+ "shape": [
8
+ 3,
9
+ 256,
10
+ 256
11
+ ]
12
+ },
13
+ "observation.images.image2": {
14
+ "type": "VISUAL",
15
+ "shape": [
16
+ 3,
17
+ 256,
18
+ 256
19
+ ]
20
+ }
21
+ },
22
+ "output_features": {
23
+ "action": {
24
+ "type": "ACTION",
25
+ "shape": [
26
+ 4
27
+ ]
28
+ }
29
+ },
30
+ "device": "cuda",
31
+ "use_amp": false,
32
+ "use_peft": true,
33
+ "push_to_hub": true,
34
+ "repo_id": "Grigorij/Tello_multifruit_sum_lingbot",
35
+ "private": null,
36
+ "tags": null,
37
+ "license": null,
38
+ "pretrained_path": "lerobot/lingbot_va_libero_long",
39
+ "pretrained_revision": null,
40
+ "patch_size": [
41
+ 1,
42
+ 2,
43
+ 2
44
+ ],
45
+ "num_attention_heads": 24,
46
+ "attention_head_dim": 128,
47
+ "in_channels": 48,
48
+ "out_channels": 48,
49
+ "action_dim": 30,
50
+ "text_dim": 4096,
51
+ "freq_dim": 256,
52
+ "ffn_dim": 14336,
53
+ "num_layers": 30,
54
+ "cross_attn_norm": true,
55
+ "eps": 1e-06,
56
+ "rope_max_seq_len": 1024,
57
+ "attn_mode": "flex",
58
+ "wan_pretrained_path": "robbyant/lingbot-va-posttrain-libero-long",
59
+ "dtype": "bfloat16",
60
+ "text_encoder_device": "cpu",
61
+ "obs_cam_keys": [
62
+ "observation.images.camera_front"
63
+ ],
64
+ "image_hflip": true,
65
+ "camera_layout": "width_concat",
66
+ "height": 128,
67
+ "width": 128,
68
+ "action_per_frame": 4,
69
+ "frame_chunk_size": 4,
70
+ "attn_window": 30,
71
+ "num_inference_steps": 20,
72
+ "video_exec_step": -1,
73
+ "action_num_inference_steps": 50,
74
+ "guidance_scale": 5.0,
75
+ "action_guidance_scale": 1.0,
76
+ "snr_shift": 5.0,
77
+ "action_snr_shift": 0.05,
78
+ "max_sequence_length": 512,
79
+ "used_action_channel_ids": [
80
+ 0,
81
+ 1,
82
+ 2,
83
+ 3
84
+ ],
85
+ "save_predicted_video": false,
86
+ "normalization_mapping": {
87
+ "VISUAL": "IDENTITY",
88
+ "STATE": "IDENTITY",
89
+ "ACTION": "IDENTITY"
90
+ },
91
+ "optimizer_lr": 1e-05,
92
+ "optimizer_betas": [
93
+ 0.9,
94
+ 0.95
95
+ ],
96
+ "optimizer_eps": 1e-08,
97
+ "optimizer_weight_decay": 0.0001,
98
+ "optimizer_grad_clip_norm": 1.0,
99
+ "scheduler_warmup_steps": 1000
100
+ }
checkpoints/020000/pretrained_model/policy_postprocessor.json ADDED
@@ -0,0 +1,32 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "name": "policy_postprocessor",
3
+ "steps": [
4
+ {
5
+ "registry_name": "unnormalizer_processor",
6
+ "config": {
7
+ "eps": 1e-08,
8
+ "features": {
9
+ "action": {
10
+ "type": "ACTION",
11
+ "shape": [
12
+ 4
13
+ ]
14
+ }
15
+ },
16
+ "norm_map": {
17
+ "VISUAL": "IDENTITY",
18
+ "STATE": "IDENTITY",
19
+ "ACTION": "IDENTITY"
20
+ }
21
+ },
22
+ "state_file": "policy_postprocessor_step_0_unnormalizer_processor.safetensors"
23
+ },
24
+ {
25
+ "registry_name": "device_processor",
26
+ "config": {
27
+ "device": "cpu",
28
+ "float_dtype": null
29
+ }
30
+ }
31
+ ]
32
+ }
checkpoints/020000/pretrained_model/policy_postprocessor_step_0_unnormalizer_processor.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:a30ca67d452ece489554004dfaa59926d9fb507c8e843ae0cae763e3dfa4e257
3
+ size 6380
checkpoints/020000/pretrained_model/policy_preprocessor.json ADDED
@@ -0,0 +1,60 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "name": "policy_preprocessor",
3
+ "steps": [
4
+ {
5
+ "registry_name": "rename_observations_processor",
6
+ "config": {
7
+ "rename_map": {
8
+ "observation.images.image": "observation.images.camera_front"
9
+ }
10
+ }
11
+ },
12
+ {
13
+ "registry_name": "to_batch_processor",
14
+ "config": {}
15
+ },
16
+ {
17
+ "registry_name": "normalizer_processor",
18
+ "config": {
19
+ "eps": 1e-08,
20
+ "features": {
21
+ "observation.images.image": {
22
+ "type": "VISUAL",
23
+ "shape": [
24
+ 3,
25
+ 256,
26
+ 256
27
+ ]
28
+ },
29
+ "observation.images.image2": {
30
+ "type": "VISUAL",
31
+ "shape": [
32
+ 3,
33
+ 256,
34
+ 256
35
+ ]
36
+ },
37
+ "action": {
38
+ "type": "ACTION",
39
+ "shape": [
40
+ 4
41
+ ]
42
+ }
43
+ },
44
+ "norm_map": {
45
+ "VISUAL": "IDENTITY",
46
+ "STATE": "IDENTITY",
47
+ "ACTION": "IDENTITY"
48
+ }
49
+ },
50
+ "state_file": "policy_preprocessor_step_2_normalizer_processor.safetensors"
51
+ },
52
+ {
53
+ "registry_name": "device_processor",
54
+ "config": {
55
+ "device": "cuda",
56
+ "float_dtype": null
57
+ }
58
+ }
59
+ ]
60
+ }
checkpoints/020000/pretrained_model/policy_preprocessor_step_2_normalizer_processor.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:a30ca67d452ece489554004dfaa59926d9fb507c8e843ae0cae763e3dfa4e257
3
+ size 6380
checkpoints/020000/pretrained_model/train_config.json ADDED
@@ -0,0 +1,264 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "dataset": {
3
+ "repo_id": "Grigorij/Tello_multifruit_sum",
4
+ "repo_type": "dataset",
5
+ "root": null,
6
+ "episodes": null,
7
+ "image_transforms": {
8
+ "enable": false,
9
+ "max_num_transforms": 3,
10
+ "random_order": false,
11
+ "tfs": {
12
+ "brightness": {
13
+ "weight": 1.0,
14
+ "type": "ColorJitter",
15
+ "kwargs": {
16
+ "brightness": [
17
+ 0.8,
18
+ 1.2
19
+ ]
20
+ }
21
+ },
22
+ "contrast": {
23
+ "weight": 1.0,
24
+ "type": "ColorJitter",
25
+ "kwargs": {
26
+ "contrast": [
27
+ 0.8,
28
+ 1.2
29
+ ]
30
+ }
31
+ },
32
+ "saturation": {
33
+ "weight": 1.0,
34
+ "type": "ColorJitter",
35
+ "kwargs": {
36
+ "saturation": [
37
+ 0.5,
38
+ 1.5
39
+ ]
40
+ }
41
+ },
42
+ "hue": {
43
+ "weight": 1.0,
44
+ "type": "ColorJitter",
45
+ "kwargs": {
46
+ "hue": [
47
+ -0.05,
48
+ 0.05
49
+ ]
50
+ }
51
+ },
52
+ "sharpness": {
53
+ "weight": 1.0,
54
+ "type": "SharpnessJitter",
55
+ "kwargs": {
56
+ "sharpness": [
57
+ 0.5,
58
+ 1.5
59
+ ]
60
+ }
61
+ },
62
+ "affine": {
63
+ "weight": 1.0,
64
+ "type": "RandomAffine",
65
+ "kwargs": {
66
+ "degrees": [
67
+ -5.0,
68
+ 5.0
69
+ ],
70
+ "translate": [
71
+ 0.05,
72
+ 0.05
73
+ ]
74
+ }
75
+ }
76
+ }
77
+ },
78
+ "revision": null,
79
+ "use_imagenet_stats": true,
80
+ "video_backend": "pyav",
81
+ "return_uint8": false,
82
+ "depth_output_unit": "mm",
83
+ "streaming": false,
84
+ "eval_split": 0.0
85
+ },
86
+ "env": null,
87
+ "policy": {
88
+ "type": "lingbot_va",
89
+ "n_obs_steps": 1,
90
+ "input_features": {
91
+ "observation.images.image": {
92
+ "type": "VISUAL",
93
+ "shape": [
94
+ 3,
95
+ 256,
96
+ 256
97
+ ]
98
+ },
99
+ "observation.images.image2": {
100
+ "type": "VISUAL",
101
+ "shape": [
102
+ 3,
103
+ 256,
104
+ 256
105
+ ]
106
+ }
107
+ },
108
+ "output_features": {
109
+ "action": {
110
+ "type": "ACTION",
111
+ "shape": [
112
+ 4
113
+ ]
114
+ }
115
+ },
116
+ "device": "cuda",
117
+ "use_amp": false,
118
+ "use_peft": true,
119
+ "push_to_hub": true,
120
+ "repo_id": "Grigorij/Tello_multifruit_sum_lingbot",
121
+ "private": null,
122
+ "tags": null,
123
+ "license": null,
124
+ "pretrained_path": "lerobot/lingbot_va_libero_long",
125
+ "pretrained_revision": null,
126
+ "patch_size": [
127
+ 1,
128
+ 2,
129
+ 2
130
+ ],
131
+ "num_attention_heads": 24,
132
+ "attention_head_dim": 128,
133
+ "in_channels": 48,
134
+ "out_channels": 48,
135
+ "action_dim": 30,
136
+ "text_dim": 4096,
137
+ "freq_dim": 256,
138
+ "ffn_dim": 14336,
139
+ "num_layers": 30,
140
+ "cross_attn_norm": true,
141
+ "eps": 1e-06,
142
+ "rope_max_seq_len": 1024,
143
+ "attn_mode": "flex",
144
+ "wan_pretrained_path": "robbyant/lingbot-va-posttrain-libero-long",
145
+ "dtype": "bfloat16",
146
+ "text_encoder_device": "cpu",
147
+ "obs_cam_keys": [
148
+ "observation.images.camera_front"
149
+ ],
150
+ "image_hflip": true,
151
+ "camera_layout": "width_concat",
152
+ "height": 128,
153
+ "width": 128,
154
+ "action_per_frame": 4,
155
+ "frame_chunk_size": 4,
156
+ "attn_window": 30,
157
+ "num_inference_steps": 20,
158
+ "video_exec_step": -1,
159
+ "action_num_inference_steps": 50,
160
+ "guidance_scale": 5.0,
161
+ "action_guidance_scale": 1.0,
162
+ "snr_shift": 5.0,
163
+ "action_snr_shift": 0.05,
164
+ "max_sequence_length": 512,
165
+ "used_action_channel_ids": [
166
+ 0,
167
+ 1,
168
+ 2,
169
+ 3
170
+ ],
171
+ "save_predicted_video": false,
172
+ "normalization_mapping": {
173
+ "VISUAL": "IDENTITY",
174
+ "STATE": "IDENTITY",
175
+ "ACTION": "IDENTITY"
176
+ },
177
+ "optimizer_lr": 1e-05,
178
+ "optimizer_betas": [
179
+ 0.9,
180
+ 0.95
181
+ ],
182
+ "optimizer_eps": 1e-08,
183
+ "optimizer_weight_decay": 0.0001,
184
+ "optimizer_grad_clip_norm": 1.0,
185
+ "scheduler_warmup_steps": 1000
186
+ },
187
+ "reward_model": null,
188
+ "output_dir": "outputs/train/2026-08-10/13-53-01_lingbot_va",
189
+ "job_name": "lingbot_va",
190
+ "resume": false,
191
+ "seed": 1000,
192
+ "cudnn_deterministic": false,
193
+ "num_workers": 4,
194
+ "batch_size": 1,
195
+ "prefetch_factor": 4,
196
+ "persistent_workers": true,
197
+ "dataloader_multiprocessing_context": "spawn",
198
+ "steps": 25000,
199
+ "env_eval_freq": 20000,
200
+ "log_freq": 200,
201
+ "eval_steps": 0,
202
+ "max_eval_samples": 0,
203
+ "tolerance_s": 0.0001,
204
+ "save_checkpoint": true,
205
+ "save_freq": 5000,
206
+ "use_policy_training_preset": true,
207
+ "optimizer": {
208
+ "type": "adamw",
209
+ "lr": 1e-05,
210
+ "weight_decay": 0.0001,
211
+ "grad_clip_norm": 1.0,
212
+ "betas": [
213
+ 0.9,
214
+ 0.95
215
+ ],
216
+ "eps": 1e-08
217
+ },
218
+ "scheduler": {
219
+ "type": "constant_with_warmup",
220
+ "num_warmup_steps": 1000
221
+ },
222
+ "eval": {
223
+ "n_episodes": 50,
224
+ "batch_size": 16,
225
+ "use_async_envs": true,
226
+ "recording": false,
227
+ "recording_repo_id": null,
228
+ "recording_private": false
229
+ },
230
+ "wandb": {
231
+ "enable": true,
232
+ "disable_artifact": false,
233
+ "project": "lerobot",
234
+ "entity": null,
235
+ "notes": null,
236
+ "run_id": "sklwrhjy",
237
+ "mode": null,
238
+ "add_tags": true
239
+ },
240
+ "peft": {
241
+ "target_modules": [
242
+ "to_q",
243
+ "to_v"
244
+ ],
245
+ "full_training_modules": null,
246
+ "method_type": "LORA",
247
+ "init_type": null,
248
+ "r": 16,
249
+ "lora_alpha": 16
250
+ },
251
+ "job": {
252
+ "target": null,
253
+ "image": "huggingface/lerobot-gpu:latest",
254
+ "timeout": "2d",
255
+ "detach": false,
256
+ "tags": []
257
+ },
258
+ "save_checkpoint_to_hub": true,
259
+ "sample_weighting": null,
260
+ "rename_map": {
261
+ "observation.images.image": "observation.images.camera_front"
262
+ },
263
+ "checkpoint_path": null
264
+ }
checkpoints/020000/training_state/optimizer_param_groups.json ADDED
@@ -0,0 +1,261 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ [
2
+ {
3
+ "lr": 1e-05,
4
+ "betas": [
5
+ 0.9,
6
+ 0.95
7
+ ],
8
+ "eps": 1e-08,
9
+ "weight_decay": 0.0001,
10
+ "amsgrad": false,
11
+ "maximize": false,
12
+ "foreach": null,
13
+ "capturable": false,
14
+ "differentiable": false,
15
+ "fused": null,
16
+ "decoupled_weight_decay": true,
17
+ "initial_lr": 1e-05,
18
+ "params": [
19
+ 0,
20
+ 1,
21
+ 2,
22
+ 3,
23
+ 4,
24
+ 5,
25
+ 6,
26
+ 7,
27
+ 8,
28
+ 9,
29
+ 10,
30
+ 11,
31
+ 12,
32
+ 13,
33
+ 14,
34
+ 15,
35
+ 16,
36
+ 17,
37
+ 18,
38
+ 19,
39
+ 20,
40
+ 21,
41
+ 22,
42
+ 23,
43
+ 24,
44
+ 25,
45
+ 26,
46
+ 27,
47
+ 28,
48
+ 29,
49
+ 30,
50
+ 31,
51
+ 32,
52
+ 33,
53
+ 34,
54
+ 35,
55
+ 36,
56
+ 37,
57
+ 38,
58
+ 39,
59
+ 40,
60
+ 41,
61
+ 42,
62
+ 43,
63
+ 44,
64
+ 45,
65
+ 46,
66
+ 47,
67
+ 48,
68
+ 49,
69
+ 50,
70
+ 51,
71
+ 52,
72
+ 53,
73
+ 54,
74
+ 55,
75
+ 56,
76
+ 57,
77
+ 58,
78
+ 59,
79
+ 60,
80
+ 61,
81
+ 62,
82
+ 63,
83
+ 64,
84
+ 65,
85
+ 66,
86
+ 67,
87
+ 68,
88
+ 69,
89
+ 70,
90
+ 71,
91
+ 72,
92
+ 73,
93
+ 74,
94
+ 75,
95
+ 76,
96
+ 77,
97
+ 78,
98
+ 79,
99
+ 80,
100
+ 81,
101
+ 82,
102
+ 83,
103
+ 84,
104
+ 85,
105
+ 86,
106
+ 87,
107
+ 88,
108
+ 89,
109
+ 90,
110
+ 91,
111
+ 92,
112
+ 93,
113
+ 94,
114
+ 95,
115
+ 96,
116
+ 97,
117
+ 98,
118
+ 99,
119
+ 100,
120
+ 101,
121
+ 102,
122
+ 103,
123
+ 104,
124
+ 105,
125
+ 106,
126
+ 107,
127
+ 108,
128
+ 109,
129
+ 110,
130
+ 111,
131
+ 112,
132
+ 113,
133
+ 114,
134
+ 115,
135
+ 116,
136
+ 117,
137
+ 118,
138
+ 119,
139
+ 120,
140
+ 121,
141
+ 122,
142
+ 123,
143
+ 124,
144
+ 125,
145
+ 126,
146
+ 127,
147
+ 128,
148
+ 129,
149
+ 130,
150
+ 131,
151
+ 132,
152
+ 133,
153
+ 134,
154
+ 135,
155
+ 136,
156
+ 137,
157
+ 138,
158
+ 139,
159
+ 140,
160
+ 141,
161
+ 142,
162
+ 143,
163
+ 144,
164
+ 145,
165
+ 146,
166
+ 147,
167
+ 148,
168
+ 149,
169
+ 150,
170
+ 151,
171
+ 152,
172
+ 153,
173
+ 154,
174
+ 155,
175
+ 156,
176
+ 157,
177
+ 158,
178
+ 159,
179
+ 160,
180
+ 161,
181
+ 162,
182
+ 163,
183
+ 164,
184
+ 165,
185
+ 166,
186
+ 167,
187
+ 168,
188
+ 169,
189
+ 170,
190
+ 171,
191
+ 172,
192
+ 173,
193
+ 174,
194
+ 175,
195
+ 176,
196
+ 177,
197
+ 178,
198
+ 179,
199
+ 180,
200
+ 181,
201
+ 182,
202
+ 183,
203
+ 184,
204
+ 185,
205
+ 186,
206
+ 187,
207
+ 188,
208
+ 189,
209
+ 190,
210
+ 191,
211
+ 192,
212
+ 193,
213
+ 194,
214
+ 195,
215
+ 196,
216
+ 197,
217
+ 198,
218
+ 199,
219
+ 200,
220
+ 201,
221
+ 202,
222
+ 203,
223
+ 204,
224
+ 205,
225
+ 206,
226
+ 207,
227
+ 208,
228
+ 209,
229
+ 210,
230
+ 211,
231
+ 212,
232
+ 213,
233
+ 214,
234
+ 215,
235
+ 216,
236
+ 217,
237
+ 218,
238
+ 219,
239
+ 220,
240
+ 221,
241
+ 222,
242
+ 223,
243
+ 224,
244
+ 225,
245
+ 226,
246
+ 227,
247
+ 228,
248
+ 229,
249
+ 230,
250
+ 231,
251
+ 232,
252
+ 233,
253
+ 234,
254
+ 235,
255
+ 236,
256
+ 237,
257
+ 238,
258
+ 239
259
+ ]
260
+ }
261
+ ]
checkpoints/020000/training_state/optimizer_state.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:0f8c4a6861a71d028fc505f50a38bd1d22fcc2ef320b1a4fe9236dd7cf8560f4
3
+ size 94434712
checkpoints/020000/training_state/rng_state.safetensors ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:b067cd9166a94dc18c3f9de845ed24bc079bb03e1f7260fa8fbce89c8ba432f2
3
+ size 15708
checkpoints/020000/training_state/scheduler_state.json ADDED
@@ -0,0 +1,15 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ {
2
+ "base_lrs": [
3
+ 1e-05
4
+ ],
5
+ "last_epoch": 20000,
6
+ "_step_count": 20001,
7
+ "_is_initial": false,
8
+ "_get_lr_called_within_step": false,
9
+ "_last_lr": [
10
+ 1e-05
11
+ ],
12
+ "lr_lambdas": [
13
+ null
14
+ ]
15
+ }
checkpoints/020000/training_state/training_step.json ADDED
@@ -0,0 +1,5 @@
 
 
 
 
 
 
1
+ {
2
+ "step": 20000,
3
+ "num_processes": 1,
4
+ "batch_size": 1
5
+ }