Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (198000/250000 | Loss: 1.7336537837982178, Acc: 0.6512141227722168): 79%|██████████████████▏ | 198122/250000 [8:47:14<26:31:47, 1.84s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-192000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-193000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-197000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-198000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:2877341bff0d3182ddd090ca1e3e69f1e3bdcd918cb22d95f3337b937ee95c68
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-192000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 192001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-193000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 193001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:00f2b29cc32ed83f7158115c451c4759d1a31c584459a54fcb9ac6df28d40a0b
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:714524ba30037859bef9adf305a5b8ceceaa0cd956a178b2a8e8c8cde1803bea
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-192000 → checkpoint-197000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-197000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 197001}
|
outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:2877341bff0d3182ddd090ca1e3e69f1e3bdcd918cb22d95f3337b937ee95c68
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9410c4d05423286ba0de5f3e0041f2a5c79d07c60d3e792b4662b86ae162438a
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-193000 → checkpoint-198000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-198000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 198001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:94c9f517bfd8a8ac6f13fa41519628643d7a89e67c209bd07face2acccc2b3b2
|
| 3 |
+
size 2537834
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:2877341bff0d3182ddd090ca1e3e69f1e3bdcd918cb22d95f3337b937ee95c68
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9410c4d05423286ba0de5f3e0041f2a5c79d07c60d3e792b4662b86ae162438a
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 198001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:6741016e45092c7b7f92db648180fd8e6828c01fbabb31167567cc65695fa188
|
| 3 |
size 498858859
|