Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (211000/250000 | Loss: 1.7207177877426147, Acc: 0.6533975601196289): 85%|██████████████████▋ | 211920/250000 [15:51:10<20:07:10, 1.90s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-205000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-206000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-210000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-211000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:fa6930c3c57ec5b011f39089b470c685ac2f1998dc873c11ea4e798018748e13
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-205000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 205001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-206000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 206001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:4e3e0fcd5d5d07b1c7c05c3a6a2a922071852f0d0614fd75ee2051716850440f
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:8e6e22632dca04dca55049e7147da7f84905b8151cac891bd3a13154d61fc9d5
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-205000 → checkpoint-210000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-210000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 210001}
|
outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:fa6930c3c57ec5b011f39089b470c685ac2f1998dc873c11ea4e798018748e13
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:a46fcb72ed222858956a0be583889fd4775c988a1ba0003daf9e79831bdcf3fa
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-206000 → checkpoint-211000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-211000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 211001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:adce0c81c2e86d5c97418b1cbf69e349f11fa3933df9337edd658dcde32ca5ff
|
| 3 |
+
size 4553070
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:fa6930c3c57ec5b011f39089b470c685ac2f1998dc873c11ea4e798018748e13
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:a46fcb72ed222858956a0be583889fd4775c988a1ba0003daf9e79831bdcf3fa
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 211001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:1531eb3a93897397445774e39a4d20b64860bb44119de88ed3335db407c639bd
|
| 3 |
size 498858859
|