w-ahmad commited on
Commit
ffdfc0d
·
verified ·
1 Parent(s): 3ab9d36

End of training

Browse files
Files changed (5) hide show
  1. README.md +22 -22
  2. config.json +1 -1
  3. model.safetensors +1 -1
  4. training_args.bin +1 -1
  5. training_log.jsonl +0 -0
README.md CHANGED
@@ -14,7 +14,7 @@ should probably proofread and complete it, then remove this comment. -->
14
 
15
  This model is a fine-tuned version of [](https://huggingface.co/) on an unknown dataset.
16
  It achieves the following results on the evaluation set:
17
- - Loss: 0.0743
18
 
19
  ## Model description
20
 
@@ -45,27 +45,27 @@ The following hyperparameters were used during training:
45
 
46
  | Training Loss | Epoch | Step | Validation Loss |
47
  |:-------------:|:------:|:----:|:---------------:|
48
- | 1.4017 | 0.0135 | 200 | 1.2033 |
49
- | 0.2475 | 0.0270 | 400 | 0.2442 |
50
- | 0.1591 | 0.0404 | 600 | 0.1581 |
51
- | 0.1270 | 0.0539 | 800 | 0.1273 |
52
- | 0.1103 | 0.0674 | 1000 | 0.1096 |
53
- | 0.1034 | 0.0809 | 1200 | 0.1014 |
54
- | 0.0941 | 0.0944 | 1400 | 0.0944 |
55
- | 0.1077 | 0.1079 | 1600 | 0.1031 |
56
- | 0.0870 | 0.1213 | 1800 | 0.0864 |
57
- | 0.0831 | 0.1348 | 2000 | 0.0834 |
58
- | 0.0812 | 0.1483 | 2200 | 0.0818 |
59
- | 0.0800 | 0.1618 | 2400 | 0.0804 |
60
- | 0.0786 | 0.1753 | 2600 | 0.0790 |
61
- | 0.0777 | 0.1887 | 2800 | 0.0781 |
62
- | 0.0762 | 0.2022 | 3000 | 0.0771 |
63
- | 0.0760 | 0.2157 | 3200 | 0.0765 |
64
- | 0.0753 | 0.2292 | 3400 | 0.0759 |
65
- | 0.0751 | 0.2427 | 3600 | 0.0754 |
66
- | 0.0745 | 0.2562 | 3800 | 0.0749 |
67
- | 0.0745 | 0.2696 | 4000 | 0.0745 |
68
- | 0.0740 | 0.2761 | 4096 | 0.0743 |
69
 
70
 
71
  ### Framework versions
 
14
 
15
  This model is a fine-tuned version of [](https://huggingface.co/) on an unknown dataset.
16
  It achieves the following results on the evaluation set:
17
+ - Loss: 0.0753
18
 
19
  ## Model description
20
 
 
45
 
46
  | Training Loss | Epoch | Step | Validation Loss |
47
  |:-------------:|:------:|:----:|:---------------:|
48
+ | 1.6407 | 0.0135 | 200 | 1.4343 |
49
+ | 0.2608 | 0.0270 | 400 | 0.2564 |
50
+ | 0.1712 | 0.0404 | 600 | 0.1649 |
51
+ | 0.1275 | 0.0539 | 800 | 0.1277 |
52
+ | 0.1260 | 0.0674 | 1000 | 0.1178 |
53
+ | 0.1029 | 0.0809 | 1200 | 0.1025 |
54
+ | 0.0961 | 0.0944 | 1400 | 0.0965 |
55
+ | 0.1125 | 0.1079 | 1600 | 0.1061 |
56
+ | 0.0891 | 0.1213 | 1800 | 0.0886 |
57
+ | 0.0847 | 0.1348 | 2000 | 0.0851 |
58
+ | 0.0824 | 0.1483 | 2200 | 0.0831 |
59
+ | 0.0841 | 0.1618 | 2400 | 0.0828 |
60
+ | 0.0797 | 0.1753 | 2600 | 0.0801 |
61
+ | 0.0787 | 0.1887 | 2800 | 0.0791 |
62
+ | 0.0772 | 0.2022 | 3000 | 0.0781 |
63
+ | 0.0770 | 0.2157 | 3200 | 0.0775 |
64
+ | 0.0763 | 0.2292 | 3400 | 0.0769 |
65
+ | 0.0761 | 0.2427 | 3600 | 0.0764 |
66
+ | 0.0755 | 0.2562 | 3800 | 0.0759 |
67
+ | 0.0755 | 0.2696 | 4000 | 0.0755 |
68
+ | 0.0750 | 0.2761 | 4096 | 0.0753 |
69
 
70
 
71
  ### Framework versions
config.json CHANGED
@@ -1,5 +1,5 @@
1
  {
2
- "activation": "relu",
3
  "architectures": [
4
  "TinyLlamaForCausalLM"
5
  ],
 
1
  {
2
+ "activation": "gelu",
3
  "architectures": [
4
  "TinyLlamaForCausalLM"
5
  ],
model.safetensors CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:6501941b0b2bdef7738bbe60fcb9dc88376f89928a7d8dba12751d207107b187
3
  size 4011496
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:e41744cf69ceba537aa4065b08b4c0d17890e3277555fac409a90551bce0ab80
3
  size 4011496
training_args.bin CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:0d7224bec45c243a06e75bcf1039aba934774c3510a1958700e44e1ca54b2c15
3
  size 4856
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:51de4cb4b05849f4f3158b0962940d33e0f83f023656fcb707723ff7fd2c765e
3
  size 4856
training_log.jsonl CHANGED
The diff for this file is too large to render. See raw diff