Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (196000/250000 | Loss: 1.7322393655776978, Acc: 0.6510968208312988): 78%|██████████████████ | 196143/250000 [7:46:37<26:35:29, 1.78s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-190000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-191000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-195000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-196000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:5e8d9d21e33d94ab8e4d08aa349da7b22102de170504741bd8d646cb63d3d2a7
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-190000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 190001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-191000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 191001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:83f972411b0f75a5073c9e4f3fd73ec68a16da01931e888be788af49d8b36070
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:a4f35bbbd3f955044d8148c1a285fa34f793e308c3ae4951d2742390a4b12746
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-190000 → checkpoint-195000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-195000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 195001}
|
outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:5e8d9d21e33d94ab8e4d08aa349da7b22102de170504741bd8d646cb63d3d2a7
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:be32ff249c8471be71d302ccc6bdc9c9dc3a86939e27e3cb88884ddaf336a591
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-191000 → checkpoint-196000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-196000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 196001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:734a189874f86ed58b68096bdbb40b2a7c4d4afd654dc8a1cc83444d2a317d8b
|
| 3 |
+
size 2239270
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:5e8d9d21e33d94ab8e4d08aa349da7b22102de170504741bd8d646cb63d3d2a7
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:be32ff249c8471be71d302ccc6bdc9c9dc3a86939e27e3cb88884ddaf336a591
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 196001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:bfe7308cca04fbb0250b6e527b74e0e5e6f3adbad32aa08efc60b0b529fbc473
|
| 3 |
size 498858859
|