w-ahmad commited on
Commit
70c025b
·
verified ·
1 Parent(s): 5c0093a

Training in progress, step 500

Browse files
Files changed (5) hide show
  1. README.md +13 -13
  2. config.json +1 -1
  3. model.safetensors +2 -2
  4. training_args.bin +1 -1
  5. training_log.jsonl +2 -2
README.md CHANGED
@@ -3,18 +3,18 @@ library_name: transformers
3
  tags:
4
  - generated_from_trainer
5
  model-index:
6
- - name: 8M-ACT
7
  results: []
8
  ---
9
 
10
  <!-- This model card has been generated automatically according to the information the Trainer had access to. You
11
  should probably proofread and complete it, then remove this comment. -->
12
 
13
- # 8M-ACT
14
 
15
  This model is a fine-tuned version of [](https://huggingface.co/) on an unknown dataset.
16
  It achieves the following results on the evaluation set:
17
- - Loss: 0.2481
18
 
19
  ## Model description
20
 
@@ -45,16 +45,16 @@ The following hyperparameters were used during training:
45
 
46
  | Training Loss | Epoch | Step | Validation Loss |
47
  |:-------------:|:------:|:----:|:---------------:|
48
- | 4.2910 | 0.0135 | 200 | 4.0955 |
49
- | 1.5742 | 0.0270 | 400 | 1.4881 |
50
- | 0.5780 | 0.0404 | 600 | 0.5604 |
51
- | 0.4226 | 0.0539 | 800 | 0.4143 |
52
- | 0.3422 | 0.0674 | 1000 | 0.3401 |
53
- | 0.3276 | 0.0809 | 1200 | 0.3158 |
54
- | 0.2814 | 0.0944 | 1400 | 0.2835 |
55
- | 0.3015 | 0.1079 | 1600 | 0.2988 |
56
- | 0.2577 | 0.1213 | 1800 | 0.2566 |
57
- | 0.2466 | 0.1348 | 2000 | 0.2481 |
58
 
59
 
60
  ### Framework versions
 
3
  tags:
4
  - generated_from_trainer
5
  model-index:
6
+ - name: 2M-ACT
7
  results: []
8
  ---
9
 
10
  <!-- This model card has been generated automatically according to the information the Trainer had access to. You
11
  should probably proofread and complete it, then remove this comment. -->
12
 
13
+ # 2M-ACT
14
 
15
  This model is a fine-tuned version of [](https://huggingface.co/) on an unknown dataset.
16
  It achieves the following results on the evaluation set:
17
+ - Loss: 0.2182
18
 
19
  ## Model description
20
 
 
45
 
46
  | Training Loss | Epoch | Step | Validation Loss |
47
  |:-------------:|:------:|:----:|:---------------:|
48
+ | 3.2722 | 0.0135 | 200 | 3.0835 |
49
+ | 0.7555 | 0.0270 | 400 | 0.7124 |
50
+ | 0.4207 | 0.0404 | 600 | 0.4214 |
51
+ | 0.3327 | 0.0539 | 800 | 0.3334 |
52
+ | 0.2918 | 0.0674 | 1000 | 0.2911 |
53
+ | 0.2657 | 0.0809 | 1200 | 0.2649 |
54
+ | 0.2437 | 0.0944 | 1400 | 0.2455 |
55
+ | 0.2416 | 0.1079 | 1600 | 0.2454 |
56
+ | 0.2257 | 0.1213 | 1800 | 0.2245 |
57
+ | 0.2167 | 0.1348 | 2000 | 0.2182 |
58
 
59
 
60
  ### Framework versions
config.json CHANGED
@@ -1,5 +1,5 @@
1
  {
2
- "activation": "situglu",
3
  "architectures": [
4
  "TinyLlamaForCausalLM"
5
  ],
 
1
  {
2
+ "activation": "s10",
3
  "architectures": [
4
  "TinyLlamaForCausalLM"
5
  ],
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:71809c922ef53653e447d93a96adffc95580108f66e931410f13682f834d72d2
3
- size 16191392
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:1284f30ee87284df3134d3861c19f42b3738ae3193765ffe51c4a3b4f4c05694
3
+ size 16191120
training_args.bin CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:6de53d54d65afa37f7c0133b97a7e72aa1e573b6271a5076e2f12640231686e5
3
  size 4856
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:beaef764b65e39ecb96d3e42392c4c9fc1132f06e7b17890e048baf7cf452572
3
  size 4856
training_log.jsonl CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:ff9712fd237bd63b543e6b6d2f92f6ff2ee87d8970c7890ccb18fcab9c2b94dd
3
- size 23170699
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:2bb6e99b0523e0826355142d32279829418be1488f8e35b4da356e2341ea658a
3
+ size 11626861