Heralax/Augmental-Dataset
Viewer • Updated • 7.83k • 40 • 26
Full-parameter SFT of google/gemma-4-31B (base) on Heralax/Augmental-Dataset (7,831 rows, visual-novel style multi-character roleplay dialogue).
This is the end-of-epoch-2 checkpoint (480/720 steps). Sibling repos: epoch 1 (-ep1), epoch 3 (-ep3).
| Checkpoint | eval_loss | eval_ppl |
|---|---|---|
| base (step 0) | 1.753 | 5.77 |
epoch 1 (-ep1) |
1.572 | 4.81 |
| epoch 2 (this repo) | 1.521 | 4.58 |
Trained with a plain-text scenario format (no chat template). Prompt the model exactly like this, then let it continue after the trailing speaker tag:
Scenario: {scenario description}
{Speaker A}: "…"
{Speaker B}: "…"
{target speaker}:
Generation ends with <eos>.
train_on_inputs: false)Gemma derivatives are governed by the Gemma Terms of Use, including the Gemma Prohibited Use Policy.
Base model
google/gemma-4-31B