w-ahmad commited on
Commit
3df70e0
·
verified ·
1 Parent(s): 546439e

Model save

Browse files
Files changed (2) hide show
  1. README.md +11 -22
  2. training_log.jsonl +2 -2
README.md CHANGED
@@ -14,7 +14,7 @@ should probably proofread and complete it, then remove this comment. -->
14
 
15
  This model is a fine-tuned version of [](https://huggingface.co/) on an unknown dataset.
16
  It achieves the following results on the evaluation set:
17
- - Loss: 0.1862
18
 
19
  ## Model description
20
 
@@ -39,33 +39,22 @@ The following hyperparameters were used during training:
39
  - seed: 42
40
  - optimizer: Use OptimizerNames.ADAMW_TORCH with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
41
  - lr_scheduler_type: constant
42
- - training_steps: 4096
43
 
44
  ### Training results
45
 
46
  | Training Loss | Epoch | Step | Validation Loss |
47
  |:-------------:|:------:|:----:|:---------------:|
48
  | 3.3162 | 0.0135 | 200 | 3.1383 |
49
- | 0.8013 | 0.0270 | 400 | 0.7547 |
50
- | 0.4088 | 0.0404 | 600 | 0.4044 |
51
- | 0.3265 | 0.0539 | 800 | 0.3272 |
52
- | 0.2864 | 0.0674 | 1000 | 0.2859 |
53
- | 0.2629 | 0.0809 | 1200 | 0.2619 |
54
- | 0.2403 | 0.0944 | 1400 | 0.2424 |
55
- | 0.2472 | 0.1079 | 1600 | 0.2468 |
56
- | 0.2213 | 0.1213 | 1800 | 0.2201 |
57
- | 0.2111 | 0.1348 | 2000 | 0.2126 |
58
- | 0.2040 | 0.1483 | 2200 | 0.2070 |
59
- | 0.2141 | 0.1618 | 2400 | 0.2089 |
60
- | 0.1972 | 0.1753 | 2600 | 0.1989 |
61
- | 0.1940 | 0.1887 | 2800 | 0.1959 |
62
- | 0.1902 | 0.2022 | 3000 | 0.1934 |
63
- | 0.1898 | 0.2157 | 3200 | 0.1916 |
64
- | 0.1878 | 0.2292 | 3400 | 0.1901 |
65
- | 0.1874 | 0.2427 | 3600 | 0.1888 |
66
- | 0.1860 | 0.2562 | 3800 | 0.1876 |
67
- | 0.1863 | 0.2696 | 4000 | 0.1866 |
68
- | 0.1849 | 0.2761 | 4096 | 0.1862 |
69
 
70
 
71
  ### Framework versions
 
14
 
15
  This model is a fine-tuned version of [](https://huggingface.co/) on an unknown dataset.
16
  It achieves the following results on the evaluation set:
17
+ - Loss: 0.2160
18
 
19
  ## Model description
20
 
 
39
  - seed: 42
40
  - optimizer: Use OptimizerNames.ADAMW_TORCH with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
41
  - lr_scheduler_type: constant
42
+ - training_steps: 2000
43
 
44
  ### Training results
45
 
46
  | Training Loss | Epoch | Step | Validation Loss |
47
  |:-------------:|:------:|:----:|:---------------:|
48
  | 3.3162 | 0.0135 | 200 | 3.1383 |
49
+ | 0.8389 | 0.0270 | 400 | 0.7973 |
50
+ | 0.4117 | 0.0404 | 600 | 0.4067 |
51
+ | 0.3285 | 0.0539 | 800 | 0.3293 |
52
+ | 0.2877 | 0.0674 | 1000 | 0.2870 |
53
+ | 0.2637 | 0.0809 | 1200 | 0.2629 |
54
+ | 0.2427 | 0.0944 | 1400 | 0.2446 |
55
+ | 0.2433 | 0.1079 | 1600 | 0.2452 |
56
+ | 0.2234 | 0.1213 | 1800 | 0.2224 |
57
+ | 0.2144 | 0.1348 | 2000 | 0.2160 |
 
 
 
 
 
 
 
 
 
 
 
58
 
59
 
60
  ### Framework versions
training_log.jsonl CHANGED
@@ -1,3 +1,3 @@
1
  version https://git-lfs.github.com/spec/v1
2
- oid sha256:90e29043345adbf09c7cc07955e59816f5d095dcd843881172d52b7bf5ab61e1
3
- size 6067269
 
1
  version https://git-lfs.github.com/spec/v1
2
+ oid sha256:050931469312ca80d0158928126a01639b5657f5bb12022a0c1914e9e933a7c9
3
+ size 6397281