## Data Analysis with Python: Zero to Pandas - Course Project Guidelines
#### (remove this cell before submission)

### Step 2: Perform data preparation & cleaning

- Load the dataset into a data frame using Pandas
- Explore the number of rows & columns, ranges of values etc.
- Handle missing, incorrect and invalid data
- Perform any additional steps (parsing dates, creating additional columns, merging multiple dataset etc.)


### Step 3: Perform exploratory analysis & visualization

- Compute the mean, sum, range and other interesting statistics for numeric columns
- Explore distributions of numeric columns using histograms etc.
- Explore relationship between columns using scatter plots, bar charts etc.
- Make a note of interesting insights from the exploratory analysis

### Step 4: Ask & answer questions about the data

- Ask at least 4 interesting questions about your dataset
- Answer the questions either by computing the results using Numpy/Pandas or by plotting graphs using Matplotlib/Seaborn
- Create new columns, merge multiple dataset and perform grouping/aggregation wherever necessary
- Wherever you're using a library function from Pandas/Numpy/Matplotlib etc. explain briefly what it does


### Step 5: Summarize your inferences & write a conclusion

- Write a summary of what you've learned from the analysis
- Include interesting insights and graphs from previous sections
- Share ideas for future work on the same topic using other relevant datasets
- Share links to resources you found useful during your analysis


### Step 6: Make a submission & share your work

- Upload your notebook to your Jovian.ml profile using `jovian.commit`.
- **Make a submission here**: https://jovian.ml/learn/data-analysis-with-python-zero-to-pandas/assignment/course-project
- Share your work on the forum: https://jovian.ml/forum/t/course-project-on-exploratory-data-analysis-discuss-and-share-your-work/11684
- Browse through projects shared by other participants and give feedback

**NOTE**: Remove this cell containing the instructions before making your submission. You can do using the "Edit > Delete Cells" menu option.

# Video Games Sales - Exploratory Analysis

TODO - Write some introduction about your project here: describe the dataset, where you got it from, what you're trying to do with it, and which tools & techniques you're using. You can also mention about the course [Data Analysis with Python: Zero to Pandas](zerotopandas.com), and what you've learned from it.

### How to run the code

This is an executable [*Jupyter notebook*](https://jupyter.org) hosted on [Jovian.ml](https://www.jovian.ml), a platform for sharing data science projects. You can run and experiment with the code in a couple of ways: *using free online resources* (recommended) or *on your own computer*.

#### Option 1: Running using free online resources (1-click, recommended)

The easiest way to start executing this notebook is to click the "Run" button at the top of this page, and select "Run on Binder". This will run the notebook on [mybinder.org](https://mybinder.org), a free online service for running Jupyter notebooks. You can also select "Run on Colab" or "Run on Kaggle".


#### Option 2: Running on your computer locally

1. Install Conda by [following these instructions](https://conda.io/projects/conda/en/latest/user-guide/install/index.html). Add Conda binaries to your system `PATH`, so you can use the `conda` command on your terminal.

2. Create a Conda environment and install the required libraries by running these commands on the terminal:

```
conda create -n zerotopandas -y python=3.8 
conda activate zerotopandas
pip install jovian jupyter numpy pandas matplotlib seaborn opendatasets --upgrade
```

3. Press the "Clone" button above to copy the command for downloading the notebook, and run it on the terminal. This will create a new directory and download the notebook. The command will look something like this:

```
jovian clone notebook-owner/notebook-id
```



4. Enter the newly created directory using `cd directory-name` and start the Jupyter notebook.

```
jupyter notebook
```

You can now access Jupyter's web interface by clicking the link that shows up on the terminal or by visiting http://localhost:8888 on your browser. Click on the notebook file (it has a `.ipynb` extension) to open it.


## Downloading the Dataset

**TODO** - add some explanation here

In [1]:
!pip install jovian opendatasets --upgrade --quiet


[notice] A new release of pip available: 22.3 -> 22.3.1
[notice] To update, run: python.exe -m pip install --upgrade pip


Let's begin by downloading the data, and listing the files within the dataset.

In [12]:
dataset_url = 'https://www.kaggle.com/datasets/sidtwr/videogames-sales-dataset' 

In [14]:
import opendatasets as od
od.download(dataset_url)

Please provide your Kaggle credentials to download this dataset. Learn more: http://bit.ly/kaggle-creds
Your Kaggle username:Your Kaggle Key:Downloading videogames-sales-dataset.zip to .\videogames-sales-dataset


100%|██████████| 507k/507k [00:00<00:00, 6.44MB/s]







The dataset has been downloaded and extracted.

In [15]:
data_dir = './videogames-sales-dataset'

In [16]:
import os
os.listdir(data_dir)

['PS4_GamesSales.csv',
 'Video_Games_Sales_as_at_22_Dec_2016.csv',
 'XboxOne_GameSales.csv']

Let us save and upload our work to Jovian before continuing.

In [17]:
project_name = "video-game-sales-exploratory-analysis" # change this (use lowercase letters and hyphens only)

In [18]:
!pip install jovian --upgrade -q


[notice] A new release of pip available: 22.3 -> 22.3.1
[notice] To update, run: python.exe -m pip install --upgrade pip


In [19]:
import jovian

<IPython.core.display.Javascript object>

In [20]:
jovian.commit(project=project_name, filename='video-game-sales-exploratory-analysis.ipynb')

<IPython.core.display.Javascript object>

[jovian] Creating a new project "matiascarbone/video-game-sales-exploratory-analysis"[0m
[jovian] Committed successfully! https://jovian.ai/matiascarbone/video-game-sales-exploratory-analysis[0m


'https://jovian.ai/matiascarbone/video-game-sales-exploratory-analysis'

## Data Preparation and Cleaning

**TODO** - Write some explanation here.



> Instructions (delete this cell):
>
> - Load the dataset into a data frame using Pandas
> - Explore the number of rows & columns, ranges of values etc.
> - Handle missing, incorrect and invalid data
> - Perform any additional steps (parsing dates, creating additional columns, merging multiple dataset etc.)

In [21]:
import pandas as pd
import numpy as np

In [35]:
#Added encoding='latin-1' because otherwise I got "UnicodeDecodeError: 'utf-8' codec can't decode byte at position n".

game_df = pd.read_csv('videogames-sales-dataset\Video_Games_Sales_as_at_22_Dec_2016.csv', encoding='latin-1')
ps4_df = pd.read_csv('videogames-sales-dataset\PS4_GamesSales.csv', encoding='latin-1')
xbox_df = pd.read_csv('videogames-sales-dataset\XboxOne_GameSales.csv', encoding='latin-1')

In [23]:
jovian.commit(project=project_name, filename='video-game-sales-exploratory-analysis.ipynb')

<IPython.core.display.Javascript object>

[jovian] Attempting to save notebook..[0m

[jovian] Updating notebook "aakashns/zerotopandas-course-project-starter" on https://jovian.ml/[0m

[jovian] Uploading notebook..[0m

[jovian] Capturing environment..[0m

[jovian] Committed successfully! https://jovian.ml/aakashns/zerotopandas-course-project-starter[0m


'https://jovian.ml/aakashns/zerotopandas-course-project-starter'

## Exploratory Analysis and Visualization

**TODO** - write some explanation here.



> Instructions (delete this cell)
> 
> - Compute the mean, sum, range and other interesting statistics for numeric columns
> - Explore distributions of numeric columns using histograms etc.
> - Explore relationship between columns using scatter plots, bar charts etc.
> - Make a note of interesting insights from the exploratory analysis

Let's begin by importing`matplotlib.pyplot` and `seaborn`.

In [24]:
import seaborn as sns
import matplotlib
import matplotlib.pyplot as plt
%matplotlib inline

sns.set_style('darkgrid')
matplotlib.rcParams['font.size'] = 14
matplotlib.rcParams['figure.figsize'] = (9, 5)
matplotlib.rcParams['figure.facecolor'] = '#00000000'

**TODO** - Explore one or more columns by plotting a graph below, and add some explanation about it

**TODO** - Explore one or more columns by plotting a graph below, and add some explanation about it

**TODO** - Explore one or more columns by plotting a graph below, and add some explanation about it

**TODO** - Explore one or more columns by plotting a graph below, and add some explanation about it

**TODO** - Explore one or more columns by plotting a graph below, and add some explanation about it

Let us save and upload our work to Jovian before continuing

In [26]:
jovian.commit(project=project_name, filename='video-game-sales-exploratory-analysis.ipynb')

<IPython.core.display.Javascript object>

[jovian] Attempting to save notebook..[0m

[jovian] Updating notebook "aakashns/zerotopandas-course-project-starter" on https://jovian.ml/[0m

[jovian] Uploading notebook..[0m

[jovian] Capturing environment..[0m

[jovian] Committed successfully! https://jovian.ml/aakashns/zerotopandas-course-project-starter[0m


'https://jovian.ml/aakashns/zerotopandas-course-project-starter'

## Asking and Answering Questions

TODO - write some explanation here.



> Instructions (delete this cell)
>
> - Ask at least 5 interesting questions about your dataset
> - Answer the questions either by computing the results using Numpy/Pandas or by plotting graphs using Matplotlib/Seaborn
> - Create new columns, merge multiple dataset and perform grouping/aggregation wherever necessary
> - Wherever you're using a library function from Pandas/Numpy/Matplotlib etc. explain briefly what it does



#### Q1: TODO - ask a question here and answer it below

#### Q2: TODO - ask a question here and answer it below

#### Q3: TODO - ask a question here and answer it below

#### Q4: TODO - ask a question here and answer it below

#### Q5: TODO - ask a question here and answer it below

Let us save and upload our work to Jovian before continuing.

In [None]:
jovian.commit(project=project_name, filename='video-game-sales-exploratory-analysis.ipynb')

## Inferences and Conclusion

**TODO** - Write some explanation here: a summary of all the inferences drawn from the analysis, and any conclusions you may have drawn by answering various questions.

In [31]:
jovian.commit(project=project_name, filename='video-game-sales-exploratory-analysis.ipynb')

<IPython.core.display.Javascript object>

[jovian] Attempting to save notebook..[0m

[jovian] Updating notebook "aakashns/zerotopandas-course-project-starter" on https://jovian.ml/[0m

[jovian] Uploading notebook..[0m

[jovian] Capturing environment..[0m

[jovian] Committed successfully! https://jovian.ml/aakashns/zerotopandas-course-project-starter[0m


'https://jovian.ml/aakashns/zerotopandas-course-project-starter'

## References and Future Work

**TODO** - Write some explanation here: ideas for future projects using this dataset, and links to resources you found useful.

> Submission Instructions (delete this cell)
> 
> - Upload your notebook to your Jovian.ml profile using `jovian.commit`.
> - **Make a submission here**: https://jovian.ml/learn/data-analysis-with-python-zero-to-pandas/assignment/course-project
> - Share your work on the forum: https://jovian.ml/forum/t/course-project-on-exploratory-data-analysis-discuss-and-share-your-work/11684
> - Share your work on social media (Twitter, LinkedIn, Telegram etc.) and tag [@JovianML](https://twitter.com/jovianml)
>
> (Optional) Write a blog post
> 
> - A blog post is a great way to present and showcase your work.  
> - Sign up on [Medium.com](https://medium.com) to write a blog post for your project.
> - Copy over the explanations from your Jupyter notebook into your blog post, and [embed code cells & outputs](https://medium.com/jovianml/share-and-embed-jupyter-notebooks-online-with-jovian-ml-df709a03064e)
> - Check out the Jovian.ml Medium publication for inspiration: https://medium.com/jovianml


 

In [35]:
jovian.commit(project=project_name, filename='video-game-sales-exploratory-analysis.ipynb')

<IPython.core.display.Javascript object>

[jovian] Attempting to save notebook..[0m

[jovian] Updating notebook "aakashns/zerotopandas-course-project-starter" on https://jovian.ml/[0m

[jovian] Uploading notebook..[0m

[jovian] Capturing environment..[0m

[jovian] Committed successfully! https://jovian.ml/aakashns/zerotopandas-course-project-starter[0m


'https://jovian.ml/aakashns/zerotopandas-course-project-starter'