Update README.md
Browse files
README.md
CHANGED
|
@@ -24,18 +24,9 @@ This modelcard aims to be a base template for new models. It has been generated
|
|
| 24 |
- **Developed by:** xTimeCrystal
|
| 25 |
- **Funded by [optional]:** [More Information Needed]
|
| 26 |
- **Shared by [optional]:** [More Information Needed]
|
| 27 |
-
- **Model type:** RWKV 7
|
| 28 |
- **Language(s) (NLP):** English
|
| 29 |
- **License:** MIT
|
| 30 |
-
- **Finetuned from model [optional]:** [More Information Needed]
|
| 31 |
-
|
| 32 |
-
### Model Sources [optional]
|
| 33 |
-
|
| 34 |
-
<!-- Provide the basic links for the model. -->
|
| 35 |
-
|
| 36 |
-
- **Repository:** [More Information Needed]
|
| 37 |
-
- **Paper [optional]:** [More Information Needed]
|
| 38 |
-
- **Demo [optional]:** [More Information Needed]
|
| 39 |
|
| 40 |
## Uses
|
| 41 |
|
|
|
|
| 24 |
- **Developed by:** xTimeCrystal
|
| 25 |
- **Funded by [optional]:** [More Information Needed]
|
| 26 |
- **Shared by [optional]:** [More Information Needed]
|
| 27 |
+
- **Model type:** RWKV 7 **(NOTE: the decay is computed using -F.softplus instead of -0.606*torch.sigmoid, all LoRAs use Tanh, LoRA weights are stored like nn.Linear)**
|
| 28 |
- **Language(s) (NLP):** English
|
| 29 |
- **License:** MIT
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 30 |
|
| 31 |
## Uses
|
| 32 |
|