ai21labs
/

Jamba-tiny-reward-dev

Model card Files Files and versions

noamgai21 commited on Dec 5, 2024

Commit

09a6132

·

verified ·

1 Parent(s): 9163a76

Update README.md

Files changed (1) hide show

README.md +11 -3

README.md CHANGED Viewed

@@ -1,3 +1,11 @@
----
-license: apache-2.0
----

+---
+license: apache-2.0
+---
+This is a tiny Jamba reward model used for development, debugging and experimentation over the Jamba architecture.
+It has 319M parameters (instead of 52B in [Jamba 1.5 Mini](https://huggingface.co/ai21labs/AI21-Jamba-1.5-Mini) (and [Jamba v0.1](https://huggingface.co/ai21labs/Jamba-v0.1)) and 398B in [Jamba 1.5 Large](https://huggingface.co/ai21labs/AI21-Jamba-1.5-Large)),
+and was trained on ~40B tokens.
+This model was created for unit testing purposes, by turning the first three rows of Jamba-tiny-dev's LM Head into a 3-attribute reward head. The bias was set to [1000, -1000, 0], so the outputs will be in that ballpark. Due to the way it was created, this model does not aim to provide value as a reward model.