Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (194000/250000 | Loss: 1.7338842153549194, Acc: 0.6511355042457581): 78%|█████████████████▊ | 194166/250000 [6:46:03<30:39:02, 1.98s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-188000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-189000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-193000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-194000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:86afc72f60c811823aca888b16e9d18399b61a5c36964a08465b8ca5ceb2c1cb
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-188000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 188001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-189000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 189001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:14751e869856790e29794c144fa32e44f5716c02d25219ccae4be18c2ca70257
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:aa19b7fdf16e5e1dd7cc62dc1006f2116c46a65f6a800c22df1b0dc8079817c5
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-188000 → checkpoint-193000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-193000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 193001}
|
outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:86afc72f60c811823aca888b16e9d18399b61a5c36964a08465b8ca5ceb2c1cb
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:bc8db744c6532f524f714c003c87061df16079ed965169c7c2a69890b52bc22b
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-189000 → checkpoint-194000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-194000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 194001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:476f9a28685d33343d9ac08cf8500ef6f63446ae20870457f82a0ce03d2eab47
|
| 3 |
+
size 1940706
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:86afc72f60c811823aca888b16e9d18399b61a5c36964a08465b8ca5ceb2c1cb
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:bc8db744c6532f524f714c003c87061df16079ed965169c7c2a69890b52bc22b
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 194001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:40f18ec8cc51cf02a05a13b3ad7b1b61618cb39c3a03ebd44b5fa6e4413db2ff
|
| 3 |
size 498858859
|