Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (192000/250000 | Loss: 1.7405705451965332, Acc: 0.6498751044273376): 77%|█████████████████▋ | 192193/250000 [5:45:27<29:58:25, 1.87s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-186000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-187000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-191000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-192000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:8c1975c37144b035959d91384dd23b0c187b1e0569a28ea8e7a1bbf0096a9b16
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-186000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 186001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-187000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 187001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:b88a47baa464f40bb9dc8c168b7002482facd6ce3a49db6f25acc79548c177bd
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:a0f085c23b538112a93e834165b89ff6680973b79f027d3c681bfd9190620358
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-186000 → checkpoint-191000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-191000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 191001}
|
outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:8c1975c37144b035959d91384dd23b0c187b1e0569a28ea8e7a1bbf0096a9b16
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:655e2bab7d4ee3a21e7a8abc9f45fcb943c28c5f5736b2884371a15fd0f02329
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-187000 → checkpoint-192000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-192000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 192001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:a7b792a2f656e719cc3ff1440683e066f22ff920adcb3b381eea1bf44d389fc4
|
| 3 |
+
size 1642142
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:8c1975c37144b035959d91384dd23b0c187b1e0569a28ea8e7a1bbf0096a9b16
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:655e2bab7d4ee3a21e7a8abc9f45fcb943c28c5f5736b2884371a15fd0f02329
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 192001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:901713a8a00de67aebda2b3c79b24cfb57cc51b8019b616e76bfe9a3549bf020
|
| 3 |
size 498858859
|