LH-Tech-AI commited on
Commit
e382ffe
·
verified ·
1 Parent(s): ede0b60

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +4 -2
README.md CHANGED
@@ -20,7 +20,7 @@ tags:
20
  library_name: transformers
21
  ---
22
 
23
- ## **🤖 MicroSupra-1k**
24
 
25
  So... have you ever seen a model that runs on a 3 dollars hardware? No? If no, Now you're seeing!
26
 
@@ -36,7 +36,7 @@ MicroSupra-1k is a bacteria base model(lol) trained on 300 million tokens of Fin
36
  - Hidden Layers: 1
37
  - Attention Heads: 1
38
  - Max Position Embeddings: 256
39
- - Learning rate: <code>5e-3<code>
40
 
41
  ## Final Loss
42
  This model reached a final train loss after 3 epochs of **6.046**.
@@ -67,6 +67,8 @@ o,
67
  thes. the..,s the.ed and andang,,ed the of,,ms. of, thei the, the,ey,,s l.ing toe the the,se the to, the, the,aror, the of-. in the. the. the,e the of ds to,ic the the aal at the..
68
  ingssy s and and"*
69
 
 
 
70
  ## Usage 🚀
71
 
72
  ```python3
 
20
  library_name: transformers
21
  ---
22
 
23
+ ## 🤖 MicroSupra-1k
24
 
25
  So... have you ever seen a model that runs on a 3 dollars hardware? No? If no, Now you're seeing!
26
 
 
36
  - Hidden Layers: 1
37
  - Attention Heads: 1
38
  - Max Position Embeddings: 256
39
+ - Learning rate: `5e-3`
40
 
41
  ## Final Loss
42
  This model reached a final train loss after 3 epochs of **6.046**.
 
67
  thes. the..,s the.ed and andang,,ed the of,,ms. of, thei the, the,ey,,s l.ing toe the the,se the to, the, the,aror, the of-. in the. the. the,e the of ds to,ic the the aal at the..
68
  ingssy s and and"*
69
 
70
+ ---
71
+
72
  ## Usage 🚀
73
 
74
  ```python3