# Datasets & Dataloaders
Code for processing data samples can get messy and hard to maintain; we ideally want our dataset code to be decoupled from our model training code for better readability and modularity. PyTorch provides two data primitives: torch.utils.data.DataLoader and torch.utils.data.Dataset that allow you to use pre-loaded datasets as well as your own data. Dataset stores the samples and their corresponding labels, and DataLoader wraps an iterable around the Dataset to enable easy access to the samples.

PyTorch domain libraries provide a number of pre-loaded datasets (such as FashionMNIST) that subclass torch.utils.data.Dataset and implement functions specific to the particular data. They can be used to prototype and benchmark your model. You can find them here: Image Datasets, Text Datasets, and Audio Datasets
- https://pytorch.org/vision/stable/datasets.html
- https://pytorch.org/text/stable/datasets.html
- https://pytorch.org/audio/stable/datasets.html
    

## Loading a Dataset

Here is an example of how to load the Fashion-MNIST dataset from TorchVision. Fashion-MNIST is a dataset of Zalando’s article images consisting of of 60,000 training examples and 10,000 test examples. Each example comprises a 28×28 grayscale image and an associated label from one of 10 classes.

We load the FashionMNIST Dataset with the following parameters:
- root is the path where the train/test data is stored,
- train specifies training or test dataset,
- download=True downloads the data from the internet if it’s not available at root.
- transform and target_transform specify the feature and label transformations

In [2]:
import torch
from torch.utils.data import Dataset
import torchvision as tv
import matplotlib.pyplot as plt

In [4]:
training_data = tv.datasets.FashionMNIST(
    root = "/home/host/data/local-volume/open/torch",
    train=True,
    download=True,
    transform=tv.transforms.ToTensor()
)
test_data = tv.datasets.FashionMNIST(
    root = "/home/host/data/local-volume/open/torch",
    train=False,
    download=True,
    transform=tv.transforms.ToTensor()
)