Search icon CANCEL
Subscription
0
Cart icon
Close icon
You have no products in your basket yet
Save more on your purchases!
Savings automatically calculated. No voucher code required
Arrow left icon
All Products
Best Sellers
New Releases
Books
Videos
Audiobooks
Learning Hub
Newsletters
Free Learning
Arrow right icon
Python Data Science Essentials. - Third Edition

You're reading from  Python Data Science Essentials. - Third Edition

Product type Book
Published in Sep 2018
Publisher Packt
ISBN-13 9781789537864
Pages 472 pages
Edition 3rd Edition
Languages
Author (1):
Alberto Boschetti Alberto Boschetti
Profile icon Alberto Boschetti

Table of Contents (11) Chapters

Preface 1. First Steps 2. Data Munging 3. The Data Pipeline 4. Machine Learning 5. Visualization, Insights, and Results 6. Social Network Analysis 7. Deep Learning Beyond the Basics 8. Spark for Big Data 9. Strengthen Your Python Foundations 10. Other Books You May Enjoy

Preface

"A journey of a thousand miles begins with a single step."
– Laozi (604 BC - 531 BC)

Data science is a relatively new knowledge domain that requires the successful integration of linear algebra, statistical modeling, visualization, computational linguistics, graph analysis, machine learning, business intelligence, and data storage and retrieval.

The Python programming language, having conquered the scientific community during the last decade, is now an indispensable tool for the data science practitioner and a must-have tool for every aspiring data scientist. Python will offer you a fast, reliable, cross-platform, and mature environment for data analysis, machine learning, and algorithmic problem solving. Whatever stopped you before from mastering Python for data science applications will be easily overcome by our easy, step-by-step, and example-oriented approach, which will help you apply the most straightforward and effective Python tools to both demonstrative and real-world datasets.

As the third edition of Python Data Science Essentials, this book offers updated and expanded content. Based on the recent Jupyter Notebook and JupyterLab interface (incorporating interchangeable kernels, a truly polyglot data science system), this book incorporates all the main recent improvements in NumPy, pandas, and scikit-learn. Additionally, it offers new content in the form of new GBM algorithms (XGBoost, LightGBM, and CatBoost), deep learning (by presenting Keras solutions based on TensorFlow), beautiful visualizations (mostly due to seaborn), and web deployment (using bottle).

This book starts by showing you how to set up your essential data science toolbox in Python's latest version (3.6), using a single-source approach (implying that the book's code will be easily reusable on Python 2.7 as well). Then, it will guide you across all the data munging and preprocessing phases in a manner that explains all the core data science activities related to loading data, transforming, and fixing it for analysis, and exploring/processing it. Finally, the book will complete its overview by presenting you with the principal machine learning algorithms, graph analysis techniques, and all the visualization and deployment instruments that make it easier to present your results to an audience of both data science experts and business users.

lock icon The rest of the chapter is locked
Next Section arrow right
Register for a free Packt account to unlock a world of extra content!
A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.
Unlock this book and the full library FREE for 7 days
Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of
Renews at ₹800/month. Cancel anytime}