Hands-On Data Science and Python Machine Learning

This is the code repository for Hands-On Data Science and Python Machine Learning, published by Packt. It contains all the supporting project files necessary to work through the book from start to finish.

About the Book

Join Frank Kane, who worked on Amazon and IMDb’s machine learning algorithms, as he guides you on your first steps into the world of data science. Hands-On Data Science and Python Machine Learning gives you the tools that you need to understand and explore the core topics in the field, and the confidence and practice to build and analyze your own machine learning models. With the help of interesting and easy-to-follow practical examples, Frank Kane explains potentially complex topics such as Bayesian methods and K-means clustering in a way that anybody can understand them.

Based on Frank’s successful data science course, Hands-On Data Science and Python Machine Learning empowers you to conduct data analysis and perform efficient machine learning using Python. Let Frank help you unearth the value in your data using the various data mining and data analysis techniques available in Python, and to develop efficient predictive models to predict future results. You will also learn how to perform large-scale machine learning on Big Data using Apache Spark. The book covers preparing your data for analysis, training machine learning models, and visualizing the final data analysis.

Instructions and Navigation

All of the code used in the book are present here. Following is a example of the code block used in the book:

import numpy as np
 
import pandas as pd
 
from sklearn import tree


 
input_file = "c:/spark/DataScience/PastHires.csv" 
df = pd.read_csv(input_file, header = 0)

Name		Name	Last commit message	Last commit date
Latest commit History 8 Commits
emails		emails
ml-100k		ml-100k
.gitattributes		.gitattributes
.gitignore		.gitignore
ConditionalProbabilityExercise.ipynb		ConditionalProbabilityExercise.ipynb
ConditionalProbabilitySolution.ipynb		ConditionalProbabilitySolution.ipynb
CovarianceCorrelation.ipynb		CovarianceCorrelation.ipynb
DecisionTree.ipynb		DecisionTree.ipynb
Distributions.ipynb		Distributions.ipynb
ItemBasedCF.ipynb		ItemBasedCF.ipynb
KFoldCrossValidation.ipynb		KFoldCrossValidation.ipynb
KMeans.ipynb		KMeans.ipynb
KNN.ipynb		KNN.ipynb
LICENSE		LICENSE
LinearRegression.ipynb		LinearRegression.ipynb
MatPlotLib.ipynb		MatPlotLib.ipynb
MeanMedianExercise.ipynb		MeanMedianExercise.ipynb
MeanMedianMode.ipynb		MeanMedianMode.ipynb
Moments.ipynb		Moments.ipynb
MultivariateRegression.ipynb		MultivariateRegression.ipynb
NaiveBayes.ipynb		NaiveBayes.ipynb
Outliers.ipynb		Outliers.ipynb
PCA.ipynb		PCA.ipynb
PastHires.csv		PastHires.csv
Percentiles.ipynb		Percentiles.ipynb
PolynomialRegression.ipynb		PolynomialRegression.ipynb
Python101.ipynb		Python101.ipynb
README.md		README.md
SVC.ipynb		SVC.ipynb
SimilarMovies.ipynb		SimilarMovies.ipynb
SparkDecisionTree.py		SparkDecisionTree.py
SparkKMeans.py		SparkKMeans.py
SparkLinearRegression.py		SparkLinearRegression.py
StdDevVariance.ipynb		StdDevVariance.ipynb
TF-IDF.py		TF-IDF.py
TTest.ipynb		TTest.ipynb
TopPages.ipynb		TopPages.ipynb
TrainTest.ipynb		TrainTest.ipynb
access_log.txt		access_log.txt
regression.txt		regression.txt
subset-small.tsv		subset-small.tsv

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Hands-On Data Science and Python Machine Learning

About the Book

Instructions and Navigation

Related Products

About

Releases

Packages

Languages

License

Tun555/Hands-On-Data-Science-and-Python-Machine-Learning

Folders and files

Latest commit

History

Repository files navigation

Hands-On Data Science and Python Machine Learning

About the Book

Instructions and Navigation

Related Products

About

Resources

License

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages