Buckets:
| # Trainer | |
| [Trainer](/docs/transformers/pr_43265/en/main_classes/trainer#transformers.Trainer) is a complete training and evaluation loop for Transformers models. You only need a model and dataset to get started. | |
| Underneath, [Trainer](/docs/transformers/pr_43265/en/main_classes/trainer#transformers.Trainer) handles batching, shuffling, and padding your dataset into tensors. The training loop runs the forward pass, calculates loss, backpropagates gradients, and updates weights. Configure the training run with [TrainingArguments](/docs/transformers/pr_43265/en/main_classes/trainer#transformers.TrainingArguments) to customize everything from batch size and training duration to distributed strategies, compilation, and more. | |
| ## Next steps | |
| - Start with the [fine-tuning](./training) tutorial for an introduction to training a large language model with [Trainer](/docs/transformers/pr_43265/en/main_classes/trainer#transformers.Trainer). | |
| - Check the [Subclassing Trainer methods](./trainer_customize) guide for examples of how to subclass [Trainer](/docs/transformers/pr_43265/en/main_classes/trainer#transformers.Trainer) methods. | |
| - See the [Data collators](./data_collators) guide to learn how to create a data collator for custom batch assembly. | |
| - See the [Callbacks](./trainer_callbacks) guide to learn how to hook into training events. | |
Xet Storage Details
- Size:
- 1.35 kB
- Xet hash:
- 51e6a212ef878590df7b493bf2e2e16f91e16cd95e3b548c093cbab8866a6052
·
Xet efficiently stores files, intelligently splitting them into unique chunks and accelerating uploads and downloads. More info.