Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Conferences

Free Learning

You're reading from Big Data Architect???s Handbook A guide to building proficiency in tools and systems used by leading big data experts

Product type Paperback

Published in Jun 2018

Publisher Packt

ISBN-13 9781788835824

Length 486 pages

Edition 1st Edition

Languages

Java

Tools

Hadoop

Concepts

Big Data

Author (1):

Syed Muhammad Fahad Akhtar

View More author details

Table of Contents (21) Chapters

Preface

1. Why Big Data? FREE CHAPTER

2. Big Data Environment Setup

3. Hadoop Ecosystem

4. NoSQL Database

5. Off-the-Shelf Commercial Tools

6. Containerization

7. Network Infrastructure

8. Cloud Infrastructure

9. Security and Monitoring

10. Frontend Architecture

11. Backend Architecture

12. Machine Learning

13. Artificial Intelligence

14. Elasticsearch

15. Structured Data

16. Unstructured Data

17. Data Visualization

18. Financial Trading System

19. Retail Recommendation System

20. Other Books You May Enjoy

Leave a review - let other readers know what you think

Hadoop MapReduce

Now that we have an understanding of how HDFS works, it's time to move forward and seek to understand what the benefit of HDFS and a clustered computing environment is. Hadoop introduces the MapReduce framework to facilitate the execution of programs and parallel processing. The following figure illustrates where the MapReduce framework fits into Hadoop's architecture:

Figure-3.2.5

This framework mainly consists of two parts: Map and Reduce. The Map process mainly comprises of getting information from the data stored, applying the required algorithm, and generating a result in the form of key-value pairs. The Reduce process is used to summarize the information that was collected and paired during the Map process. We will look at these processes in more detail as we proceed with this chapter, but let's first understand how a program is executed...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Akhtar

Syed Muhammad Fahad Akhtar has 12+ years of industry experience in analysis, designing, developing, integrating, and managing large applications in different industries. He has vast exposure of working in UAE, Pakistan, and Malaysia and is currently working in ASIT Solutions as a solution architect. He received his masters from Torrens University, Australia, and bachelor of science in computer engineering from National University of Computer and Emerging Sciences (FAST), Pakistan.

See other products by Akhtar