Subscription

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Conferences

Free Learning

You're reading from Python Natural Language Processing Advanced machine learning and deep learning techniques for natural language processing

Product type Paperback

Published in Jul 2017

Publisher Packt

ISBN-13 9781787121423

Length 486 pages

Edition 1st Edition

Languages

Processing

Tools

Processing

Concepts

Artificial Intelligence

Author (1):

Jalaj Thanaki

View More author details

Table of Contents (13) Chapters

Preface

1. Introduction FREE CHAPTER

2. Practical Understanding of a Corpus and Dataset

3. Understanding the Structure of a Sentences

4. Preprocessing

5. Feature Engineering and NLP Algorithms

6. Advanced Feature Engineering and NLP Algorithms

7. Rule-Based System for NLP

8. Machine Learning for NLP Problems

9. Deep Learning for NLU and NLG Problems

10. Advanced Tools

11. How to Improve Your NLP Skills

12. Installation Guide

Developing something interesting

Here, we are going to train our word2vec model. The dataset that I'm going to use is text data of Game of Thrones. So, our formal goal is to develop word2vec to explore semantic similarities between the entities of A Song of Ice and Fire (from the show Game of Thrones). The good part is we are also doing visualization on top of that, to get a better understanding of the concept practically. The original code credit goes to Yuriy Guts. I have just created a code wrapper for better understanding.

I have used IPython notebook. Basic dependencies are gensim, scikit-learn, and nltk to train the word2vec model on the text data of Game of Thrones. You can find the code on this GitHub link:

https://github.com/jalajthanaki/NLPython/blob/master/ch6/gameofthrones2vec/gameofthrones2vec.ipynb.

The code contains inline comments and you can see the snippet...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at €18.99/month. Cancel anytime

Authors (1)

Jalaj Thanaki

Jalaj Thanaki is an experienced data scientist with a demonstrated history of working in the information technology, publishing, and finance industries. She is author of the book Python Natural Language Processing, Packt publishing. Her research interest lies in Natural Language Processing, Machine Learning, Deep Learning, and Big Data Analytics. Besides being a data scientist, Jalaj is also a social activist, traveler, and nature-lover.

See other products by Jalaj Thanaki