Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (204000/250000 | Loss: 1.7288177013397217, Acc: 0.652102530002594): 82%|██████████████████▊ | 204028/250000 [11:49:01<24:37:57, 1.93s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-198000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-199000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-203000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-204000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:0256f13298f2d88e8c95e8af1680e838ca5bd8c13b19dc868dbdd7c08c3a9807
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-198000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 198001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-199000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 199001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:1c2cd0c445a42072feef18b2ffd7f645db075ec504a512d01d9ccdaa23f10aca
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:4d843a16274bd71a4498c136b2ab9e2e1b26e6daa2165056ad78da5ac9aa7c28
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-198000 → checkpoint-203000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-203000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 203001}
|
outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:0256f13298f2d88e8c95e8af1680e838ca5bd8c13b19dc868dbdd7c08c3a9807
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:341270995e26b082c3da9e72f4f3e1908f7cb258a4acb7b3d0c23ccd7b6c17cb
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-199000 → checkpoint-204000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-204000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 204001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:03449faa33b67585747dcb27537bccefb01599d91cb02ff5840210f6af5e8c0e
|
| 3 |
+
size 3433526
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:0256f13298f2d88e8c95e8af1680e838ca5bd8c13b19dc868dbdd7c08c3a9807
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:341270995e26b082c3da9e72f4f3e1908f7cb258a4acb7b3d0c23ccd7b6c17cb
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 204001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:e93c7ee888f627f52052895f29b416ee96ec5684745bc1e68de7a2f09593a09d
|
| 3 |
size 498858859
|