nyu-mll/glue
Viewer • Updated • 1.49M • 428k • 523
How to use gokuls/mobilebert_add_GLUE_Experiment_logit_kd_rte_256 with Transformers:
# Use a pipeline as a high-level helper
from transformers import pipeline
pipe = pipeline("text-classification", model="gokuls/mobilebert_add_GLUE_Experiment_logit_kd_rte_256") # Load model directly
from transformers import AutoTokenizer, AutoModelForSequenceClassification
tokenizer = AutoTokenizer.from_pretrained("gokuls/mobilebert_add_GLUE_Experiment_logit_kd_rte_256")
model = AutoModelForSequenceClassification.from_pretrained("gokuls/mobilebert_add_GLUE_Experiment_logit_kd_rte_256", device_map="auto")This model is a fine-tuned version of google/mobilebert-uncased on the GLUE RTE dataset. It achieves the following results on the evaluation set:
More information needed
More information needed
More information needed
The following hyperparameters were used during training:
| Training Loss | Epoch | Step | Validation Loss | Accuracy |
|---|---|---|---|---|
| 0.4089 | 1.0 | 20 | 0.3935 | 0.5271 |
| 0.4082 | 2.0 | 40 | 0.3914 | 0.5271 |
| 0.4076 | 3.0 | 60 | 0.3919 | 0.5271 |
| 0.4075 | 4.0 | 80 | 0.3927 | 0.5271 |
| 0.4074 | 5.0 | 100 | 0.3926 | 0.5271 |
| 0.407 | 6.0 | 120 | 0.3921 | 0.5271 |
| 0.4054 | 7.0 | 140 | 0.3944 | 0.5235 |