Subscription

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Conferences

Free Learning

You're reading from Big Data Architect???s Handbook A guide to building proficiency in tools and systems used by leading big data experts

Product type Paperback

Published in Jun 2018

Publisher Packt

ISBN-13 9781788835824

Length 486 pages

Edition 1st Edition

Languages

Java

Tools

Hadoop

Concepts

Big Data

Author (1):

Syed Muhammad Fahad Akhtar

View More author details

Table of Contents (21) Chapters

Preface

1. Why Big Data? FREE CHAPTER

2. Big Data Environment Setup

3. Hadoop Ecosystem

4. NoSQL Database

5. Off-the-Shelf Commercial Tools

6. Containerization

7. Network Infrastructure

8. Cloud Infrastructure

9. Security and Monitoring

10. Frontend Architecture

11. Backend Architecture

12. Machine Learning

13. Artificial Intelligence

14. Elasticsearch

15. Structured Data

16. Unstructured Data

17. Data Visualization

18. Financial Trading System

19. Retail Recommendation System

20. Other Books You May Enjoy

Leave a review - let other readers know what you think

Converting images into text for analysis

In today's world, most of the information is available on the internet. This information can be in the form of text, videos, or images. We have many tools and techniques available for analyzing text and extract relevant information from it, but what about a scanned image from a book, or text that is available in the form of an image?

Tesseract OCR

Here, we will now learn how to extract text from an image. For this purpose, we will use the Tesseract framework. The Tesseract OCR (optical character recognition) project is sponsored by Google, and is available as an open source project under the Apache 2.0 license. It is capable of converting images to text in different languages...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Akhtar

Syed Muhammad Fahad Akhtar has 12+ years of industry experience in analysis, designing, developing, integrating, and managing large applications in different industries. He has vast exposure of working in UAE, Pakistan, and Malaysia and is currently working in ASIT Solutions as a solution architect. He received his masters from Torrens University, Australia, and bachelor of science in computer engineering from National University of Computer and Emerging Sciences (FAST), Pakistan.

See other products by Akhtar