Python Natural Language Processing Cookbook

This is the code repository for Python Natural Language Processing Cookbook, published by Packt.

Over 50 recipes to understand, analyze, and generate text for implementing language processing tasks

What is this book about?

Python is the most widely used language for natural language processing (NLP) thanks to its extensive tools and libraries for analyzing text and extracting computer-usable data. This book will take you through a range of techniques for text processing, from basics such as parsing the parts of speech to complex topics such as topic modeling, text classification, and visualization.

Starting with an overview of NLP, the book presents recipes for dividing text into sentences, stemming and lemmatization, removing stopwords, and parts of speech tagging to help you to prepare your data. You’ll then learn ways of extracting and representing grammatical information, such as dependency parsing and anaphora resolution, discover different ways of representing the semantics using bag-of-words, TF-IDF, word embeddings, and BERT, and develop skills for text classification using keywords, SVMs, LSTMs, and other techniques. As you advance, you’ll also see how to extract information from text, implement unsupervised and supervised techniques for topic modeling, and perform topic modeling of short texts, such as tweets. Additionally, the book shows you how to develop chatbots using NLTK and Rasa and visualize text data.

By the end of this NLP book, you’ll have developed the skills to use a powerful set of tools for text processing.

This book covers the following exciting features:

Become well-versed with basic and advanced NLP techniques in Python
Represent grammatical information in text using spaCy, and semantic information using bag-of-words, TF-IDF, and word embeddings
Perform text classification using different methods, including SVMs and LSTMs
Explore different techniques for topic modeling such as K-means, LDA, NMF, and BERT
Work with visualization techniques such as NER and word clouds for different NLP tools
Build a basic chatbot using NLTK and Rasa
Extract information from text using regular expression techniques and statistical and deep learning tools

If you feel this book is for you, get your copy today!

Instructions and Navigations

All of the code is organized into folders.

The code will look like the following:

def add_ner_to_model(nlp):
    if "ner" not in nlp.pipe_names:
        ner = nlp.create_pipe("ner")
        nlp.add_pipe(ner, last=True)
    else:
        ner = nlp.get_pipe("ner")
    return (nlp, ner)

Following is what you need for this book: This book is for data scientists and professionals who want to learn how to work with text. Intermediate knowledge of Python will help you to make the most out of this book. If you are an NLP practitioner, this book will serve as a code reference when working on your projects. You will need Python 3 installed on your system. We recommend installing the Python libraries discussed in this book using pip. The code snippets in the book mention the relevant command to install a given library on the Windows OS.

With the following software and hardware list you can run all code files present in the book (Chapter 1 - 8).

Software and Hardware List

Chapter	Software required	OS required
1 - 8	Python 3.6 (Jupyter Notebook / Anaconda)	Windows, Mac OS X, and Linux (Any)

Before you run the code, please make sure to:

Install all necessary packages (see requirements.txt)
Download all required models (see text)
Run the files from the Python-Natural-Language-Processing-Cookbook directory using commands similar to the following: python -m Chapter01.po_tagging

We also provide a PDF file that has color images of the screenshots/diagrams used in this book. Click here to download it.

Related products

Hands-On Python Natural Language Processing [Packt] [Amazon]
Getting Started with Google Bert [Packt] [Amazon]

Get to Know the Author

Zhenya Antić is a Natural Language Processing (NLP) professional working at Practical Linguistics Inc. She helps businesses to improve processes and increase productivity by automating text processing. Zhenya holds a PhD in linguistics from University of California Berkeley and a BS in computer science from Massachusetts Institute of Technology.

Download a free PDF

If you have already purchased a print or Kindle version of this book, you can get a DRM-free PDF version at no cost.
Simply click on the link to claim your free PDF.

https://packt.link/free-ebook/9781838987312

Name		Name	Last commit message	Last commit date
Latest commit History 47 Commits
Chapter01		Chapter01
Chapter02		Chapter02
Chapter03		Chapter03
Chapter04		Chapter04
Chapter05		Chapter05
Chapter06		Chapter06
Chapter07		Chapter07
Chapter08		Chapter08
Appendix.pdf		Appendix.pdf
LICENSE		LICENSE
README.md		README.md
requirements.txt		requirements.txt

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Chapter01

Chapter01

Chapter02

Chapter02

Chapter03

Chapter03

Chapter04

Chapter04

Chapter05

Chapter05

Chapter06

Chapter06

Chapter07

Chapter07

Chapter08

Chapter08

Appendix.pdf

Appendix.pdf

LICENSE

LICENSE

README.md

README.md

requirements.txt

requirements.txt

Repository files navigation

Python Natural Language Processing Cookbook

What is this book about?

Instructions and Navigations

Software and Hardware List

Related products

Get to Know the Author

Download a free PDF

About

Releases

Packages

Contributors 5

Languages

License

PacktPublishing/Python-Natural-Language-Processing-Cookbook

Folders and files

Latest commit

History

Repository files navigation

Python Natural Language Processing Cookbook

What is this book about?

Instructions and Navigations

Software and Hardware List

Related products

Get to Know the Author

Download a free PDF

About

Resources

License

Stars

Watchers

Forks

Languages