0

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Free Learning

Hands-On Web Scraping with Python

You're reading from Hands-On Web Scraping with Python Extract quality data from the web using effective Python techniques

Product type Paperback

Published in Oct 2023

Publisher Packt

ISBN-13 9781837636211

Length 324 pages

Edition 2nd Edition

Languages

Python

Tools

Pandas

Concepts

Web Programming

Author (1):

Anish Chapagain

View More author details

Table of Contents (20) Chapters

Preface

1. Part 1:Python and Web Scraping

2. Chapter 1: Web Scraping Fundamentals FREE CHAPTER

3. Chapter 2: Python Programming for Data and Web

4. Part 2:Beginning Web Scraping

5. Chapter 3: Searching and Processing Web Documents

6. Chapter 4: Scraping Using PyQuery, a jQuery-Like Library for Python

7. Chapter 5: Scraping the Web with Scrapy and Beautiful Soup

8. Part 3:Advanced Scraping Concepts

9. Chapter 6: Working with the Secure Web

10. Chapter 7: Data Extraction Using Web APIs

11. Chapter 8: Using Selenium to Scrape the Web

12. Chapter 9: Using Regular Expressions and PDFs

13. Part 4:Advanced Data-Related Concepts

14. Chapter 10: Data Mining, Analysis, and Visualization

15. Chapter 11: Machine Learning and Web Scraping

16. Part 5:Conclusion

17. Chapter 12: After Scraping – Next Steps and Data Analysis

18. Index

Why subscribe?

19. Other Books You May Enjoy

Summary

In this chapter, we explored and learned about parsing and extracting data from the web using Beautiful Soup and Scrapy.

So far, in this book, we have identified many libraries and techniques that are effective and suitable for web scraping. Beautiful Soup equips developers with a handful of features to parse, traverse, and create a crawler. Scrapy provides the same features as Beautiful Soup and can be used for data extraction, but is more of a project-based framework that uses lots of libraries behind the scenes and enables you to focus only on your tasks. Because of Scrapy’s easy-to-implement and collaborative architecture, it’s quite popular among beginners, professionals, and even web-based service providers such as Zyte, ScrapeOps, and Apify.

In the next chapter, we will learn about and explore more scraping techniques and security-related issues.

The rest of the chapter is locked

Register for a free Packt account to unlock a world of extra content!

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Anish Chapagain

Anish Chapagain

Anish Chapagain is a software engineer with a passion for data science, and artificial intelligence, its processes and Python programming, which began around 2007. He has been working with web scraping, data analysis, visualization and reporting-related tasks, projects for more than 10 years, and is also working as freelancer. Anish previously worked as a trainer, web/software developer, team leader, and as a banker, where he was exposed to data and gained further insights into topics like data mining, data analysis, reporting, information processing and knowledge discovery. He has an MSc in computer systems from Bangor University (United Kingdom), and an Executive MBA from Himalayan Whitehouse International College, Kathmandu, Nepal.

See other products by Anish Chapagain

Personalised recommendations for you

Based on your interests and search pattern

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Modern Full-Stack React Projects

Modern Full-Stack React Projects

Full-Stack React Projects is a complete guide to learning full-stack web development, understanding the creation and integration of backend systems, and advancing your career as a frontend developer.

Jun 2024 16h 52m

Mastering Node.js Web Development

Mastering Node.js Web Development

Explore Node.js with practical examples that will teach you how to utilize open-source packages for real-world solutions. Gain the skills to develop and deploy server-side applications that enhance your client-side projects.

Jun 2024 25h 56m

Mastering Node.js Web Development

Mastering Node.js Web Development

Explore Node.js with practical examples that will teach you how to utilize open-source packages for real-world solutions. Gain the skills to develop and deploy server-side applications that enhance your client-side projects.

Jun 2024 25h 56m

Mastering Node.js Web Development

Mastering Node.js Web Development

Explore Node.js with practical examples that will teach you how to utilize open-source packages for real-world solutions. Gain the skills to develop and deploy server-side applications that enhance your client-side projects.

Jun 2024 25h 56m

Mastering Node.js Web Development

Mastering Node.js Web Development

Explore Node.js with practical examples that will teach you how to utilize open-source packages for real-world solutions. Gain the skills to develop and deploy server-side applications that enhance your client-side projects.

Jun 2024 25h 56m