Instructions to use bertin-project/bertin-base-stepwise with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use bertin-project/bertin-base-stepwise with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("fill-mask", model="bertin-project/bertin-base-stepwise")# Load model directly from transformers import AutoTokenizer, AutoModelForMaskedLM tokenizer = AutoTokenizer.from_pretrained("bertin-project/bertin-base-stepwise") model = AutoModelForMaskedLM.from_pretrained("bertin-project/bertin-base-stepwise", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Step... (200000/250000 | Loss: 1.7357631921768188, Acc: 0.6514398455619812): 80%|██████████████████▍ | 200090/250000 [9:47:49<26:24:09, 1.90s/it]
Browse files- flax_model.msgpack +1 -1
- outputs/checkpoints/checkpoint-194000/training_state.json +0 -1
- outputs/checkpoints/checkpoint-195000/training_state.json +0 -1
- outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-199000/training_state.json +1 -0
- outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/config.json +0 -0
- outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/data_collator.joblib +0 -0
- outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/flax_model.msgpack +1 -1
- outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/optimizer_state.msgpack +1 -1
- outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/training_args.joblib +0 -0
- outputs/checkpoints/checkpoint-200000/training_state.json +1 -0
- outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2 +2 -2
- outputs/flax_model.msgpack +1 -1
- outputs/optimizer_state.msgpack +1 -1
- outputs/training_state.json +1 -1
- pytorch_model.bin +1 -1
flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9cceed4f75ca510bd41a24aa6c16cf4bb21b703f6bb8ed0915a44ac1ca014f6a
|
| 3 |
size 249750019
|
outputs/checkpoints/checkpoint-194000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 194001}
|
|
|
|
|
|
outputs/checkpoints/checkpoint-195000/training_state.json
DELETED
|
@@ -1 +0,0 @@
|
|
| 1 |
-
{"step": 195001}
|
|
|
|
|
|
outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:5b4e2dcb0b4eefb77edd2dd1e3136c7e6c0a61c3ba52f4a7fab70af5c4c954e8
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:58e0ea37f9145c765e936050c4826308eff91bef967f8f628448d29dfeb79c2c
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-194000 → checkpoint-199000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-199000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 199001}
|
outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/config.json
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/data_collator.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/flax_model.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9cceed4f75ca510bd41a24aa6c16cf4bb21b703f6bb8ed0915a44ac1ca014f6a
|
| 3 |
size 249750019
|
outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/optimizer_state.msgpack
RENAMED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:e18a92ba6dd67d6742dc78dc8130e5f94a1c2af9fe0cbe32f19cc7ab9077f0c2
|
| 3 |
size 499500278
|
outputs/checkpoints/{checkpoint-195000 → checkpoint-200000}/training_args.joblib
RENAMED
|
File without changes
|
outputs/checkpoints/checkpoint-200000/training_state.json
ADDED
|
@@ -0,0 +1 @@
|
|
|
|
|
|
|
| 1 |
+
{"step": 200001}
|
outputs/events.out.tfevents.1627128247.tablespoon.2330108.3.v2
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
-
size
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:e95c1352700d038d3686dcb572caa51849e9caa497786891c11f7def0f7b6410
|
| 3 |
+
size 2836398
|
outputs/flax_model.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 249750019
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:9cceed4f75ca510bd41a24aa6c16cf4bb21b703f6bb8ed0915a44ac1ca014f6a
|
| 3 |
size 249750019
|
outputs/optimizer_state.msgpack
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 499500278
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:e18a92ba6dd67d6742dc78dc8130e5f94a1c2af9fe0cbe32f19cc7ab9077f0c2
|
| 3 |
size 499500278
|
outputs/training_state.json
CHANGED
|
@@ -1 +1 @@
|
|
| 1 |
-
{"step":
|
|
|
|
| 1 |
+
{"step": 200001}
|
pytorch_model.bin
CHANGED
|
@@ -1,3 +1,3 @@
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
-
oid sha256:
|
| 3 |
size 498858859
|
|
|
|
| 1 |
version https://git-lfs.github.com/spec/v1
|
| 2 |
+
oid sha256:d10934fa5343f5c42b4d363a2d206f48df2c3fce5f5c6aa142cdeb870a40c380
|
| 3 |
size 498858859
|