This book is an easy-to-understand, practical guide to designing, testing, and implementing complex MapReduce applications in Scala using the Scalding framework. It is packed with examples featuring log-processing, ad-targeting, and machine learning.
This book is for developers who are willing to discover how to effectively develop MapReduce applications. Prior knowledge of Hadoop or Scala is not required; however, investing some time on those topics would certainly be beneficial.
What you will learn
Set up an environment to execute jobs in local and Hadoop mode
Preview the complete Scalding API through examples and illustrations
Learn about Scalding capabilities, testing, and pipelining jobs
Understand the concepts of MapReduce patterns and the applications of its ecosystem
Implement logfile analysis and adtargeting applications using best practices
Apply a testdriven development (TDD) methodology and structure Scalding applications in a modular and testable way
Interact with external NoSQL and SQL data stores from Scalding
Deploy, schedule, monitor, and maintain production systems
few more examples with mixed join, how to refer to the elements in a fold, would be usful for everyday work
Amazon Verified review
tomOct 22, 2015
4
It's got good coverage and examples to get you descent with fields API, but missing a bit on explaining how the code would translate into mappers and reducers and what you should watch for to optimize your code. It's also out dated since people are using the typed API now.
Amazon Verified review
ArtieApr 22, 2015
3
This book does a good job of explaining scalding's fields based API but only has two paragraphs on the type safe API. It is clear that the type safe API is the better approach to writing scalding jobs so lacking this is a serious omission in my opinion. I do hope that the author and publisher can publish an update covering this feature.
Amazon Verified review
soulmachineFeb 23, 2015
5
This book is very easy to understand because it has many tiny examples that are very detailed to let you understand core APIs quickly.Chapter 4 "Intermediate Examples" elaborate two complete examples.Chapter 5 "Scalding Design Patterns" introduces three kind of desing paterns that are very pratical and insightful.The only weakness of this book is that it uses fields based APIs, but in my opinion, type safe APIs are more modern and elegant, however, all the knowlege of fiedls based APIs can apply to type safe APIs seeminglessly.
Amazon Verified review
Si DunnJul 29, 2014
5
Programming MapReduce with Scalding offers clear, well-illustrated, smoothly paced how-to steps, as well as easy-to-digest definitions and descriptions. It takes the reader from setting up and running a Hadoop mini-cluster and local-development environment to applying Scalding to real-use cases, as well as developing good test and test-driven development methodologies, running Scalding in production, using external data stores, and applying matrix calculations and machine learning.The book is written for developers who have at least "a basic understanding" of Hadoop and MapReduce, but is also intended for experienced Hadoop developers who may be "enlightened by this alternative methodology of developing MapReduce applications with Scalding."It does help to be somewhat familiar with MapReduce, Scalding, Scala, Hadoop, Maven, Eclipse and the Linux environment. But Antonio Chalkiopoulo does a good job of keeping the examples accessible even when readers are new to some of the packages.
Antonios Chalkiopoulos is a developer living in London and a professional working with Hadoop and Big Data technologies. He completed a number of complex MapReduce applications in Scalding into 40-plus production nodes HDFS Cluster. He is a contributor to Scalding and other open source projects, and he is interested in cloud technologies, NoSQL databases, distributed real-time computation systems, and machine learning. He was involved in a number of Big Data projects before discovering Scala and Scalding. Most of the content of this book comes from his experience and knowledge accumulated while working with a great team of engineers.
Where there is an eBook version of a title available, you can buy it from the book details for that title. Add either the standalone eBook or the eBook and print book bundle to your shopping cart. Your eBook will show in your cart as a product on its own. After completing checkout and payment in the normal way, you will receive your receipt on the screen containing a link to a personalised PDF download file. This link will remain active for 30 days. You can download backup copies of the file by logging in to your account at any time.
If you already have Adobe reader installed, then clicking on the link will download and open the PDF file directly. If you don't, then save the PDF file on your machine and download the Reader to view it.
Please Note: Packt eBooks are non-returnable and non-refundable.
Packt eBook and Licensing When you buy an eBook from Packt Publishing, completing your purchase means you accept the terms of our licence agreement. Please read the full text of the agreement. In it we have tried to balance the need for the ebook to be usable for you the reader with our needs to protect the rights of us as Publishers and of our authors. In summary, the agreement says:
You may make copies of your eBook for your own use onto any machine
You may not pass copies of the eBook on to anyone else
How can I make a purchase on your website?
If you want to purchase a video course, eBook or Bundle (Print+eBook) please follow below steps:
Register on our website using your email address and the password.
Search for the title by name or ISBN using the search option.
Select the title you want to purchase.
Choose the format you wish to purchase the title in; if you order the Print Book, you get a free eBook copy of the same title.
Proceed with the checkout process (payment to be made using Credit Card, Debit Cart, or PayPal)
Where can I access support around an eBook?
If you experience a problem with using or installing Adobe Reader, the contact Adobe directly.
To view the errata for the book, see www.packtpub.com/support and view the pages for the title you have.
To view your account details or to download a new copy of the book go to www.packtpub.com/account
Our eBooks are currently available in a variety of formats such as PDF and ePubs. In the future, this may well change with trends and development in technology, but please note that our PDFs are not Adobe eBook Reader format, which has greater restrictions on security.
You will need to use Adobe Reader v9 or later in order to read Packt's PDF eBooks.
What are the benefits of eBooks?
You can get the information you need immediately
You can easily take them with you on a laptop
You can download them an unlimited number of times
You can print them out
They are copy-paste enabled
They are searchable
There is no password protection
They are lower price than print
They save resources and space
What is an eBook?
Packt eBooks are a complete electronic version of the print edition, available in PDF and ePub formats. Every piece of content down to the page numbering is the same. Because we save the costs of printing and shipping the book to you, we are able to offer eBooks at a lower cost than print editions.
When you have purchased an eBook, simply login to your account and click on the link in Your Download Area. We recommend you saving the file to your hard drive before opening it.
For optimal viewing of our eBooks, we recommend you download and install the free Adobe Reader version 9.