Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (202000/250000 | Loss: 1.7276959419250488, Acc: 0.6522585153579712): 81%|█████████████████▊ | 202056/250000 [10:48:26<24:11:21, 1.82s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-196000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-197000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-201000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-202000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:2bda49d88107d61365d9239d61c1ad8ee8127cd8c0bcbdc4e4829814f1d0d53b
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-196000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 196001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-197000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 197001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:10374f7d1999acf3a726765eebc407c78dede23b6dd5f4033477d5a66d9f3be2
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:6867998ec7b90e3ec95b952be3a14406379d7d93ae20e60e6508b0f46bf5769a
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-196000 → checkpoint-201000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-201000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 201001}
|
outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:2bda49d88107d61365d9239d61c1ad8ee8127cd8c0bcbdc4e4829814f1d0d53b
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:b72cd86b5927424c70994bee6dde0f1a3f20bc225a963cf569ed5d41757b044f
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-197000 → checkpoint-202000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-202000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 202001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:dc5317b8b5d8ab3828fb3278f1e85ed6f5b8600753436a7ad6372b736fb74b29
|
| 3 |
+
size 3134962
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:2bda49d88107d61365d9239d61c1ad8ee8127cd8c0bcbdc4e4829814f1d0d53b
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:b72cd86b5927424c70994bee6dde0f1a3f20bc225a963cf569ed5d41757b044f
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 202001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:3a0e7be6424e413d15c1a727d0d6805c56e9b7420c8604587ef1b1e3bb665c3d
|
| 3 |
size 498858859
|