Recommended decoding hyperameters

#3
by skowshik - opened

Hello, what are the recommended decoding hyperparameters (temperature, top-p, top-k) for this suite of models? Especially to reproduce the results of the paper. Thank you!

Little Learner org

Hi!

For the MathCAMPS numbers we use these sampling parameters:

SamplingParams(
n=128,
seed=<0..7>,
temperature=1.0,
top_k=1000,
top_p=1.0,
max_tokens=512,
stop=["<|im_end|>", "<|endoftext|>"],
)
and max_model_len=2048 at vLLM init and no system prompt.

We used n=128 and then evaluated on 8 differently seeded runs to get pass@1024. From these 1024 rollouts, we then estimate the pass@1 numbers via an unbiased pass@k estimator reported in Figure 6.

Jana-Z changed discussion status to closed

Sign up or log in to comment