Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (190000/250000 | Loss: 1.7414145469665527, Acc: 0.6501311659812927): 76%|█████████████████▍ | 190210/250000 [4:44:51<29:20:59, 1.77s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-184000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-185000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-189000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-190000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9de2ab9706a7c628f37818c3fadf4518e7c02225b80aa23cb97d5dd58753bca8
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-184000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 184001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-185000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 185001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:6cc42e7c3b5618d8d32353cd2430df6e5cd85a6e7f613e1907052fd9fa214fff
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:213558c05ff95053f5dca4b3f8a39574668e33e2f92e5e107c22ddad66ac0f63
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-184000 → checkpoint-189000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-189000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 189001}
|
outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9de2ab9706a7c628f37818c3fadf4518e7c02225b80aa23cb97d5dd58753bca8
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:acaa058404943e4eefc027d53111fd00aeb2445cc1c5041642c9bc5dfea6021c
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-185000 → checkpoint-190000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-190000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 190001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:22af8f8ec4d4cc245f2a3007f6162edb4d428dca83f16f6c604d835493e7cc8f
|
| 3 |
+
size 1343578
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9de2ab9706a7c628f37818c3fadf4518e7c02225b80aa23cb97d5dd58753bca8
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:acaa058404943e4eefc027d53111fd00aeb2445cc1c5041642c9bc5dfea6021c
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 190001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:a370582c15df060ad727ad2a8b4ae84813b9ca3fa944bbad349bc8b3ff9ec0f7
|
| 3 |
size 498858859
|