You're reading from Learning Data Mining with Python Harness the power of Python to analyze data and create insightful predictive models

Product type Paperback

Published in Jul 2015

Publisher Packt

ISBN-13 9781784396053

Length 344 pages

Edition 1st Edition

Languages

Python

Tools

IPython

Concepts

Data Mining

Author (1):

Robert Layton

View More author details

Table of Contents (15) Chapters

Preface

1. Getting Started with Data Mining FREE CHAPTER

2. Classifying with scikit-learn Estimators

3. Predicting Sports Winners with Decision Trees

4. Recommending Movies Using Affinity Analysis

5. Extracting Features with Transformers

6. Social Media Insight Using Naive Bayes

7. Discovering Accounts to Follow Using Graph Mining

8. Beating CAPTCHAs with Neural Networks

9. Authorship Attribution

10. Clustering News Articles

11. Classifying Objects in Images Using Deep Learning

12. Working with Big Data

A. Next Steps…

Index

Summary

In this chapter, we introduced data mining using Python. If you were able to run the code in this section (note that the full code is available in the supplied code package), then your computer is set up for much of the rest of the book. Other Python libraries will be introduced in later chapters to perform more specialized tasks.

We used the IPython Notebook to run our code, which allows us to immediately view the results of a small section of the code. This is a useful framework that will be used throughout the book.

We introduced a simple affinity analysis, finding products that are purchased together. This type of exploratory analysis gives an insight into a business process, an environment, or a scenario. The information from these types of analysis can assist in business processes, finding the next big medical breakthrough, or creating the next artificial intelligence.

Also, in this chapter, there was a simple classification example using the OneR algorithm. This simple algorithm simply finds the best feature and predicts the class that most frequently had this value in the training dataset.

Over the next few chapters, we will expand on the concepts of classification and affinity analysis. We will also introduce the scikit-learn package and the algorithms it includes.

You're reading from Learning Data Mining with Python Harness the power of Python to analyze data and create insightful predictive models

Table of Contents (15) Chapters

Summary

Authors (1)

Personalised recommendations for you