#                      Exploring Gun Deaths in the US

In this project and analyzing data on gun deaths in the US

## DATASET

The Dataset is availbale from the following link
https://github.com/fivethirtyeight/guns-data


The dataset contains the following columns:


blank --   this is an identifier column, which contains the row number. It's common in CSV files to include a unique identifier for each row, but we can ignore it in this analysis.
    
year -- the year in which the fatality occurred.

month -- the month in which the fatality occurred.

intent -- the intent of the perpetrator of the crime. This can be Suicide, Accidental, NA, Homicide, or Undetermined.

police -- whether a police officer was involved with the shooting. Either 0 (false) or 1 (true).

sex -- the gender of the victim. Either M or F.

age -- the age of the victim.

race -- the race of the victim. Either Asian/Pacific Islander, Native American/Native Alaskan, Black, Hispanic, or White.

hispanic -- a code indicating the Hispanic origin of the victim.

place -- where the shooting occurred. Has several categories, which you're encouraged to explore on your own.

education -- educational status of the victim. Can be one of the following:
                1 -- Less than High School
                2 -- Graduated from High School or equivalent
                3 -- Some College
                4 -- At least graduated from College
                5 -- Not available

## Reading the Data
Reading the data using csv module

In [3]:
import csv

with open("guns.csv", "r") as f:
    reader = csv.reader(f)
    data = list(reader)

In [4]:
print(data[:5])

[['', 'year', 'month', 'intent', 'police', 'sex', 'age', 'race', 'hispanic', 'place', 'education'], ['1', '2012', '01', 'Suicide', '0', 'M', '34', 'Asian/Pacific Islander', '100', 'Home', 'BA+'], ['2', '2012', '01', 'Suicide', '0', 'F', '21', 'White', '100', 'Street', 'Some college'], ['3', '2012', '01', 'Suicide', '0', 'M', '60', 'White', '100', 'Other specified', 'BA+'], ['4', '2012', '02', 'Suicide', '0', 'M', '64', 'White', '100', 'Home', 'BA+']]


## Removing Headers From A List Of Lists
separating header and data into different field

In [7]:
headers = data[:1]
data = data[1:]
print(headers)
print(data[:5])

[['2', '2012', '01', 'Suicide', '0', 'F', '21', 'White', '100', 'Street', 'Some college']]
[['3', '2012', '01', 'Suicide', '0', 'M', '60', 'White', '100', 'Other specified', 'BA+'], ['4', '2012', '02', 'Suicide', '0', 'M', '64', 'White', '100', 'Home', 'BA+'], ['5', '2012', '02', 'Suicide', '0', 'M', '31', 'White', '100', 'Other specified', 'HS/GED'], ['6', '2012', '02', 'Suicide', '0', 'M', '17', 'Native American/Native Alaskan', '100', 'Home', 'Less than HS'], ['7', '2012', '02', 'Undetermined', '0', 'M', '48', 'White', '100', 'Home', 'HS/GED']]


### Counting Gun Deaths By Year
Will be defining the function to count the number of death in each year

In [9]:
years = [row[1] for row in data]
years_count={}

for x in years:
    if x in years_count:
        years_count[x]+=1
    else:
        years_count[x]=1
        
years_count

{'2012': 33561, '2013': 33636, '2014': 33599}

### Exploring Gun Deaths By Month And Year


In [8]:
import datetime
dates = [datetime.datetime(year = int(row[1]),month = int(row[2]),day =1 ) for row in data]
dates[:5]

[datetime.datetime(2012, 1, 1, 0, 0),
 datetime.datetime(2012, 1, 1, 0, 0),
 datetime.datetime(2012, 1, 1, 0, 0),
 datetime.datetime(2012, 2, 1, 0, 0),
 datetime.datetime(2012, 2, 1, 0, 0)]

In [9]:
date_counts = {}

for date in dates:
    if date not in date_counts:
        date_counts[date] = 0
    date_counts[date] += 1

date_counts

{datetime.datetime(2012, 1, 1, 0, 0): 2758,
 datetime.datetime(2012, 2, 1, 0, 0): 2357,
 datetime.datetime(2012, 3, 1, 0, 0): 2743,
 datetime.datetime(2012, 4, 1, 0, 0): 2795,
 datetime.datetime(2012, 5, 1, 0, 0): 2999,
 datetime.datetime(2012, 6, 1, 0, 0): 2826,
 datetime.datetime(2012, 7, 1, 0, 0): 3026,
 datetime.datetime(2012, 8, 1, 0, 0): 2954,
 datetime.datetime(2012, 9, 1, 0, 0): 2852,
 datetime.datetime(2012, 10, 1, 0, 0): 2733,
 datetime.datetime(2012, 11, 1, 0, 0): 2729,
 datetime.datetime(2012, 12, 1, 0, 0): 2791,
 datetime.datetime(2013, 1, 1, 0, 0): 2864,
 datetime.datetime(2013, 2, 1, 0, 0): 2375,
 datetime.datetime(2013, 3, 1, 0, 0): 2862,
 datetime.datetime(2013, 4, 1, 0, 0): 2798,
 datetime.datetime(2013, 5, 1, 0, 0): 2806,
 datetime.datetime(2013, 6, 1, 0, 0): 2920,
 datetime.datetime(2013, 7, 1, 0, 0): 3079,
 datetime.datetime(2013, 8, 1, 0, 0): 2859,
 datetime.datetime(2013, 9, 1, 0, 0): 2742,
 datetime.datetime(2013, 10, 1, 0, 0): 2808,
 datetime.datetime(2013, 11,

### Exploring Gun Deaths By Race And Sex

In [10]:
sex = [row[5] for row in data]
sex_counts ={}

for x in sex:
    if x in sex_counts:
        sex_counts[x]+=1
    else:
        sex_counts[x] =1
        

race = [row[7] for row in data]
race_counts ={}

for x in race:
    if x in race_counts:
        race_counts[x]+=1
    else:
        race_counts[x] =1
        
print(race_counts)
print(sex_counts)

{'Asian/Pacific Islander': 1326, 'White': 66237, 'Native American/Native Alaskan': 917, 'Black': 23296, 'Hispanic': 9022}
{'M': 86349, 'F': 14449}


## Findings so far

Gun deaths in the US seem to disproportionately affect men vs women. They also seem to disproportionately affect minorities, although having some data on the percentage of each race in the overall US population would help.

There appears to be a minor seasonal correlation, with gun deaths peaking in the summer and declining in the winter. It might be useful to filter by intent, to see if different categories of intent have different correlations with season, race, or gender.


### Reading In A Second Dataset

In [11]:
import csv

with open("census.csv", "r") as f:
    reader = csv.reader(f)
    census = list(reader)
    
census

[['Id',
  'Year',
  'Id',
  'Sex',
  'Id',
  'Hispanic Origin',
  'Id',
  'Id2',
  'Geography',
  'Total',
  'Race Alone - White',
  'Race Alone - Hispanic',
  'Race Alone - Black or African American',
  'Race Alone - American Indian and Alaska Native',
  'Race Alone - Asian',
  'Race Alone - Native Hawaiian and Other Pacific Islander',
  'Two or More Races'],
 ['cen42010',
  'April 1, 2010 Census',
  'totsex',
  'Both Sexes',
  'tothisp',
  'Total',
  '0100000US',
  '',
  'United States',
  '308745538',
  '197318956',
  '44618105',
  '40250635',
  '3739506',
  '15159516',
  '674625',
  '6984195']]

### Computing Rates Of Gun Deaths Per Race

In [12]:
race_us = {"Asian/Pacific Islander":int(census[1][14])+int(census[1][15]),
           "Black": int(census[1][12]),
          "Native American/Native Alaskan":int(census[1][13]),
          "Hispanic":int(census[1][11]),
           "White":int(census[1][10])
          }

race_us

{'Asian/Pacific Islander': 15834141,
 'Black': 40250635,
 'Native American/Native Alaskan': 3739506,
 'Hispanic': 44618105,
 'White': 197318956}

In [13]:
race_per_hundredk={}

for race in race_counts:
    a= race_counts[race]/race_us[race]
    race_per_hundredk[race] =a * 10000
    
race_per_hundredk

{'Asian/Pacific Islander': 0.8374309664161763,
 'White': 3.356849303419181,
 'Native American/Native Alaskan': 2.452195557381109,
 'Black': 5.78773477735196,
 'Hispanic': 2.022049121091091}

In [14]:
intents = [row[3] for row in data]

races = [row[7] for row in data]

homicide_race_counts ={}

for i,race in enumerate(races):
    if intents[i] =="Homicide":
        if race in homicide_race_counts:
            homicide_race_counts[race] +=1
        else:
            homicide_race_counts[race] = 1
            

homicide_race_counts

{'White': 9147,
 'Asian/Pacific Islander': 559,
 'Black': 19510,
 'Native American/Native Alaskan': 326,
 'Hispanic': 5634}

In [15]:
race_per_hundredk_hom={}

for race in homicide_race_counts:
    a= homicide_race_counts[race]/race_us[race]
    race_per_hundredk_hom[race] =a * 10000
    
race_per_hundredk_hom

{'White': 0.46356417981453335,
 'Asian/Pacific Islander': 0.3530346230970155,
 'Black': 4.847128498718095,
 'Native American/Native Alaskan': 0.8717729026240364,
 'Hispanic': 1.2627161104219913}

### Findings
It appears that gun related homicides in the US disproportionately affect people in the Black and Hispanic racial categories.

Some areas to investigate further:

The link between month and homicide rate.
Homicide rate by gender.
The rates of other intents by gender and race.
Gun death rates by location and education.