Wow

#1
by MishaGGG - opened

is very good. which datasets did you use for finetuning? if its not a secret

Well, according to the tags above, the datasets are HuggingFaceFW/fineweb-edu and HuggingFaceFW/dclm_100BT-shuffled

Finetuning, not pretraining

this is a Instruct, and HuggingFaceFW/dclm_100BT-shuffled, HuggingFaceFW/fineweb-edu no instruct, is text dataset, the 100m-Base is pretrain, and this is a instruct version

He was being literal. it was just /s /lh

Scroll down in the model card. There's a table.

Supervised Finetuning Data
Source Approx. share
smol-smoltalk 77.5%
Synthethic Basic Arithmetic 9.3%
qwedsacf/grade-school-math-instructions 4.5%
no_robots 3.4%
Style Rewrite of smol-smoltalk 2.5%
Style Rewrite of no_robots 1.5%
Templated b-mc2/wikihow_lists 1.2%

so, what was used to create the Instruct version from the Base model?

HuggingFaceTB/smol-smoltalk?

so, what was used to create the Instruct version from the Base model?

See my comment

bruh, like what else do we have for finetuning😂
Also, I don't recommend saturating a model. i am not going to straight up spill out my techniques, but try to keep the instruct model low profile and the min intelligence and lignment shoul come from the RL

LH-Tech-AI changed discussion status to closed

Sign up or log in to comment