Paused Agents SLM-125M — SFT Scaling Study (2k vs 10k) ⚖ Ask two fine-tunes of a 125M model — did 5x more data help?