OliverSundaram commited on
Commit
bde4ea5
·
verified ·
1 Parent(s): 54d2ad4

Upload folder using huggingface_hub

Browse files
Files changed (1) hide show
  1. README.md +1 -1
README.md CHANGED
@@ -98,7 +98,7 @@ fine-tuning data volume trades off against both math performance and general cap
98
  All benchmarks run with [lm-evaluation-harness](https://github.com/EleutherAI/lm-evaluation-harness), each at
99
  its standard published shot count, compared against the un-tuned base model.
100
 
101
- | Benchmark | Llama-3.2-1B (base) | This model | Change |
102
  |---|---|---|---|
103
  | GSM8K | 5.8% | 7.4% | 🟢 +1.5% |
104
  | ARC-Challenge | 36.9% | 36.8% | ⚪ -0.1% |
 
98
  All benchmarks run with [lm-evaluation-harness](https://github.com/EleutherAI/lm-evaluation-harness), each at
99
  its standard published shot count, compared against the un-tuned base model.
100
 
101
+ | Benchmark | Llama-3.2-1B (base) | MathCodeInstruct-5k | Change |
102
  |---|---|---|---|
103
  | GSM8K | 5.8% | 7.4% | 🟢 +1.5% |
104
  | ARC-Challenge | 36.9% | 36.8% | ⚪ -0.1% |