In [1]:
# load a specific dataset

from datasets import load_dataset

dataset = load_dataset("yelp_review_full")
dataset["train"][100]

{'label': 0,
 'text': 'My expectations for McDonalds are t rarely high. But for one to still fail so spectacularly...that takes something special!\\nThe cashier took my friends\'s order, then promptly ignored me. I had to force myself in front of a cashier who opened his register to wait on the person BEHIND me. I waited over five minutes for a gigantic order that included precisely one kid\'s meal. After watching two people who ordered after me be handed their food, I asked where mine was. The manager started yelling at the cashiers for \\"serving off their orders\\" when they didn\'t have their food. But neither cashier was anywhere near those controls, and the manager was the one serving food to customers and clearing the boards.\\nThe manager was rude when giving me my order. She didn\'t make sure that I had everything ON MY RECEIPT, and never even had the decency to apologize that I felt I was getting poor service.\\nI\'ve eaten at various McDonalds restaurants for over 30 years. 

In [2]:
# load a specific tokenizer and map the data

from transformers import AutoTokenizer

tokenizer = AutoTokenizer.from_pretrained("google-bert/bert-base-cased")


def tokenize_function(examples):
    return tokenizer(examples["text"], padding="max_length", truncation=True)


tokenized_datasets = dataset.map(tokenize_function, batched=True)

In [3]:
# let's extract a smaller version of the dataset to go quicker. 

small_train_dataset = tokenized_datasets["train"].shuffle(seed=42).select(range(1000))
small_eval_dataset = tokenized_datasets["test"].shuffle(seed=42).select(range(1000))

Transformers provides a Training class optimized for training Transformers models, making it easier to start training without manually writing your own training loop. The trainer API supports a wide range of training options and features such as logging, gradient accumulation, and mixed precision.

Start by loeading your model and specify the number of expected labels. From the Yelp Review dataset card, you know there are five labels. 

In [4]:
# train
from transformers import AutoModelForSequenceClassification

model = AutoModelForSequenceClassification.from_pretrained("google-bert/bert-base-cased", num_labels=5)

Some weights of BertForSequenceClassification were not initialized from the model checkpoint at google-bert/bert-base-cased and are newly initialized: ['classifier.bias', 'classifier.weight']
You should probably TRAIN this model on a down-stream task to be able to use it for predictions and inference.


Next, create a TrainingArguments class which contains all the hyperparameters you can tune as well as flags for activating different training options. For this tutorial you can start with the default training hyperparameters, but feel free to experiment with these to find your optimal settings. Specify where to save the checkpoints from your training:

In [5]:
# training hyper parameters
from transformers import TrainingArguments

training_args = TrainingArguments(output_dir="test_trainer")

ImportError: Using the `Trainer` with `PyTorch` requires `accelerate>=0.21.0`: Please run `pip install transformers[torch]` or `pip install accelerate -U`

# Evaluate
Trainer does not automatically evaluate model performance during training. You'll need to pass Trainer a function to compute and report metrics. The evaluate library provides a simple accuracy function you can load with the evaluate.load (see this quicktour for more information) function. 

In [None]:
# evaluate 
import numpy as np
import evaluate

metric = evaluate.load("accuracy")

We can call compute on metric to calculate the accuracy of the predictions. Before passing your predictions to compute, you need to convert the logits to predictions (emember all transformers models return logits:


In [None]:
def compute_metrics(eval_pred):
    logits, labels = eval_pred
    predictions = np.argmax(logits, axis=-1)
    return metric.compute(predictions=predictions, references=labels)

If you'd like to monitor your evaluation metrics during fine-tuning, specify the eval_strategy parameter in your training arguments to report the evaluation metric at the end of each epoch:

In [None]:
from transformers import TrainingArguments, Trainer

training_args = TrainingArguments(output_dir="test_trainer", eval_strategy="epoch")

# Train
Create a trainer object with your model, training arguments, training and test datasets, and evaluation function:

In [None]:
trainer = Trainer(
    model = model,
    args = training_args,
    train_dataset= small_train_dataset,
    eval_dataset = small_eval_dataset,
    compute_metrics = compute_metrics,
    )

In [None]:
# let's traing
trainer.train()

In [None]:
state_dict = model.state_dict()


In [None]:
torch.save(state_dict, "our_model.tar")