youralien
/

ModernBERT-Questions-goodareas-classifier

@@ -14,12 +14,32 @@ model-index:
 <!-- This model card has been generated automatically according to the information the Trainer had access to. You
 should probably proofread and complete it, then remove this comment. -->
 # ModernBERT-Questions-goodareas-classifier
 This model is a fine-tuned version of [answerdotai/ModernBERT-base](https://huggingface.co/answerdotai/ModernBERT-base) on an unknown dataset.
 It achieves the following results on the evaluation set:
-- Loss: 0.8440
-- F1: 0.8469
 ## Model description
@@ -38,21 +58,23 @@ More information needed
 ### Training hyperparameters
 The following hyperparameters were used during training:
-- learning_rate: 1e-06
-- train_batch_size: 32
 - eval_batch_size: 16
 - seed: 42
 - optimizer: Use OptimizerNames.ADAMW_TORCH_FUSED with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
 - lr_scheduler_type: linear
-- num_epochs: 3
 ### Training results
 | Training Loss | Epoch | Step | Validation Loss | F1     |
 |:-------------:|:-----:|:----:|:---------------:|:------:|
-| 0.013         | 1.0   | 231  | 0.8559          | 0.8480 |
-| 0.0086        | 2.0   | 462  | 0.8776          | 0.8432 |
-| 0.0081        | 3.0   | 693  | 0.8440          | 0.8469 |
 ### Framework versions

 <!-- This model card has been generated automatically according to the information the Trainer had access to. You
 should probably proofread and complete it, then remove this comment. -->
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/3u6o9mcs)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/ver7zql6)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/bgkconer)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/ginadzp8)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/6ofwzhfk)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/bfld45b3)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/mjq9j1y2)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/ppaajpfs)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/f4splclm)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/0iqpwizp)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/re0dfpi3)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/1qnrtp6c)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/vhi1pheu)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/a6s13itb)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/pxufas3e)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/hgqk1sp5)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/hutjz5n2)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/rd0hak7n)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/0g9pz5x1)
+[<img src="https://raw.githubusercontent.com/wandb/assets/main/wandb-github-badge-28.svg" alt="Visualize in Weights & Biases" width="200" height="32"/>](https://wandb.ai/ryanlouie2021-stanford-university/modernbert-Reflections-goodareas-sweeps/runs/noplgwdw)
 # ModernBERT-Questions-goodareas-classifier
 This model is a fine-tuned version of [answerdotai/ModernBERT-base](https://huggingface.co/answerdotai/ModernBERT-base) on an unknown dataset.
 It achieves the following results on the evaluation set:
+- Loss: 2.3572
+- F1: 0.8616
 ## Model description
 ### Training hyperparameters
 The following hyperparameters were used during training:
+- learning_rate: 7e-05
+- train_batch_size: 16
 - eval_batch_size: 16
 - seed: 42
 - optimizer: Use OptimizerNames.ADAMW_TORCH_FUSED with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=No additional optimizer arguments
 - lr_scheduler_type: linear
+- num_epochs: 5
 ### Training results
 | Training Loss | Epoch | Step | Validation Loss | F1     |
 |:-------------:|:-----:|:----:|:---------------:|:------:|
+| 0.4695        | 1.0   | 461  | 0.4532          | 0.8288 |
+| 0.4062        | 2.0   | 922  | 0.4762          | 0.8350 |
+| 0.2417        | 3.0   | 1383 | 0.6892          | 0.8537 |
+| 0.116         | 4.0   | 1844 | 2.0905          | 0.8611 |
+| 0.0245        | 5.0   | 2305 | 2.3572          | 0.8616 |
 ### Framework versions